Exemplar Performance provides recipes in ready-to-use templates for evaluating performance of specific AI use cases across hardware and software combinations.
☆112Aug 26, 2026Updated 3 weeks ago
Alternatives and similar repositories for exemplar-performance
Users that are interested in exemplar-performance are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- nvloom is a set of tools designed to scalably test MNNVL fabrics.☆65Jul 24, 2026Updated last month
- NVSentinel detects and remediates GPU faults on Kubernetes nodes☆383Updated this week
- A TUI-based utility for real-time monitoring of InfiniBand traffic and performance metrics on the local node☆74May 16, 2026Updated 4 months ago
- Dynamo Workshop☆19Nov 7, 2025Updated 10 months ago
- ☆13May 30, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A toolkit for discovering cluster network topology.☆170Updated this week
- ☆53Updated this week
- ☆68Updated this week
- ☆27Oct 9, 2025Updated 11 months ago
- A comprehensive Helm chart for monitoring GPU resources in Kubernetes clusters. This tool provides real-time visibility into GPU allocati…☆34Aug 26, 2026Updated 3 weeks ago
- A benchmark framework for Pytorch☆34Mar 14, 2025Updated last year
- Guides and examples to help achieve optimal performance on a NVIDIA Grace CPU☆18Aug 9, 2024Updated 2 years ago
- NVIDIA NCCL Tests for Distributed Training☆159Updated this week
- These are lab guides for our customer/partner training and lab environment☆20Apr 15, 2026Updated 5 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Pavilion is a Python 3 (3.6+) based framework for running and analyzing tests targeting HPC systems.☆46Sep 10, 2026Updated last week
- Benchmarking guide for the Azure AI Infrastructure.☆39Jun 9, 2026Updated 3 months ago
- Prometheus collector and exporter for Slurm cluster metrics. A Slinky project.☆18Nov 7, 2025Updated 10 months ago
- NVIDIA Fleet Intelligence Agent - Host agent for GPU telemetry collection and attestation☆53Updated this week
- Run Slurm on Kubernetes. A Slinky project.☆362Updated this week
- Linux Sysinfo Snapshot☆69Jun 7, 2026Updated 3 months ago
- A Streamlit app for exploring available AWS EC2 Capacity Blocks and SageMaker Training Plans across regions and instance types.☆30Updated this week
- HPC tests using MPI codes & synthetic benchmarks with IB/RoCE comparisions - from StackHPC Ltd.☆22Jul 11, 2022Updated 4 years ago
- A Slurm-based HPC workload management environment, driven by Ansible.☆73Updated this week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Tooling for optimized, validated, and reproducible GPU-accelerated AI runtime in Kubernetes☆421Updated this week
- A distributed storage benchmark for file systems, object stores & block devices with support for GPUs☆293Updated this week
- Slurm debian packages☆16Updated this week
- NVIDIA fork of QEMU☆17Sep 9, 2026Updated last week
- Scripts to customize AWS ParallelCluster☆30Jun 11, 2026Updated 3 months ago
- A tool for bandwidth measurements on NVIDIA GPUs.☆776Jul 28, 2026Updated last month
- CloudAI Benchmark Framework☆99Updated this week
- Python wrappers for the FirecREST API☆12Sep 9, 2026Updated last week
- Prototype of OpenSHMEM for NVIDIA GPUs, developed as part of DoE Design Forward☆24Apr 26, 2018Updated 8 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- InfiniBand fabric monitoring daemon written in Go☆32May 22, 2025Updated last year
- Parallel Computing -- Validation Suite: Validation engine for Exascale project benchmarks☆16Mar 26, 2026Updated 5 months ago
- NVIDIA Infra Controller - Hardware Lifecycle Management and multitenant networking☆271Updated this week
- Recipes for reproducing training and serving benchmarks for large machine learning models using GPUs on Google Cloud.☆140Sep 8, 2026Updated last week
- DRA Driver for NVIDIA GPUs☆711Updated this week
- ATLAHS: An Application-centric Network Simulator Toolchain for AI, HPC, and Distributed Storage☆103May 12, 2026Updated 4 months ago
- ☆38Oct 31, 2025Updated 10 months ago