Slurm Exporter is a Prometheus exporter designed to scrape and expose a comprehensive range of performance and scheduling metrics from Slurm-managed clusters. It supports both CPU and GPU resource accounting, node and partition state monitoring, job tracking, and scheduler statistics.
☆50Aug 10, 2026Updated this week
Alternatives and similar repositories for slurm_exporter
Users that are interested in slurm_exporter are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A Prometheus exporter and a REST API server to export metrics of compute units of resource managers like SLURM, LSF, Openstack, k8s☆74Updated this week
- Converts an Infiniband topology file to graphviz dot format or slurm topology.conf format☆18Feb 2, 2026Updated 6 months ago
- YAML-based database of datacenter infrastructures☆33Jun 4, 2026Updated 2 months ago
- Helm chart and Terraform modules for the JARVICE XE Hybrid Cloud HPC platform☆18Updated this week
- A coherent Ansible roles collection to simply deploy clusters of nodes.☆163Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- SLURM-native software GPU slicing for NVIDIA clusters using memory limits and compute time-slicing.☆51May 24, 2026Updated 2 months ago
- A remote registry for Singularity Registry HPC 🖊️☆16Updated this week
- Slurm Lua SPANK plugin☆17Jan 30, 2025Updated last year
- Prometheus exporter for the stats in the cgroup accounting with slurm. This will also collect stats of a job using NVIDIA GPUs.☆49Jul 27, 2026Updated 2 weeks ago
- Prometheus exporter for lustre☆27Updated this week
- Slurm Exporter for Prometheus☆19May 27, 2026Updated 2 months ago
- cluster power control☆50Jun 8, 2026Updated 2 months ago
- Tool to profile usage of HPC resources by regularly probing processes.☆12Jun 24, 2026Updated last month
- Slurm-Mail is a drop in replacement for Slurm's e-mails to give users much more information about their jobs compared to the standard Slu…☆122Jul 25, 2026Updated 2 weeks ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Export select slurm metrics to prometheus☆68Feb 19, 2026Updated 5 months ago
- CLI tools for Slurm clusters☆14Jun 5, 2026Updated 2 months ago
- Identify and reduce instances of underutilization by the users of high-performance computing systems☆18Jul 13, 2026Updated 3 weeks ago
- A daemon that uses cgroups to monitor and manage user behavior on login nodes☆82Aug 3, 2026Updated last week
- A Slurm-based HPC workload management environment, driven by Ansible.☆72Updated this week
- ☆14Dec 24, 2021Updated 4 years ago
- Dump slurm accounting database to sqlite3 database for easy analysis☆19Jun 26, 2026Updated last month
- Zero Downtime Cumulus Upgrades with Ansible☆10Jul 8, 2016Updated 10 years ago
- Tool to prep huge data volumes for place in archives like Glacier, Data Den, or HPSS☆31Jul 21, 2026Updated 3 weeks ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Jobstats is a job monitoring platform for CPU and GPU clusters☆136Jul 29, 2026Updated last week
- Command-line tool to retrieve information and monitor Mellanox un-managed Infiniband switches☆77Nov 17, 2025Updated 8 months ago
- Slurm to InfluxDB stats collection script☆16Jan 17, 2022Updated 4 years ago
- s9s - Slurm 9000 Text User Interface☆39Jun 26, 2026Updated last month
- Ansible role to install/configure Pacemaker cluster resource manager☆16Mar 28, 2026Updated 4 months ago
- Pavilion is a Python 3 (3.6+) based framework for running and analyzing tests targeting HPC systems.☆46Updated this week
- Container-based Slurm cluster with support for running on multiple ssh-accessible computers. Currently it is based on podman, systemd, no…☆25Dec 21, 2020Updated 5 years ago
- Prometheus exporter for collection of RDMA (RoCE) NIC statistics from Linux hosts☆43Updated this week
- DataDog agent check for Linux Conntrack metrics☆10Feb 1, 2019Updated 7 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Information for the Intro to Cluster System Administration for Non-Sysadmins class☆10Dec 12, 2021Updated 4 years ago
- Info on CHPC Open OnDemand installation and customization☆17Jul 1, 2025Updated last year
- ☆10Jul 27, 2026Updated 2 weeks ago
- User Fencing Tools☆16Mar 2, 2022Updated 4 years ago
- Pipeline for poreathon☆14Dec 17, 2014Updated 11 years ago
- Generic Puppet Configuration for HPC Clusters☆14Sep 8, 2018Updated 7 years ago
- EmbeddedFab provides Open Source AVR (Atmega32) Library☆14Oct 22, 2015Updated 10 years ago