Automated bottleneck detection and solution orchestration
☆23Feb 24, 2026Updated 5 months ago
Alternatives and similar repositories for intelliperf
Users that are interested in intelliperf are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- IntelliKit is a collection of intelligent tools designed to make GPU kernel development, profiling, and validation accessible to LLMs and…☆31Jul 30, 2026Updated last week
- AMD RAD's multi-GPU Triton-based framework for seamless multi-GPU programming☆195Updated this week
- Scale-out system monitoring☆26Updated this week
- A lightweight, general-purpose framework for evaluating GPU kernel and benchmark.☆77Updated this week
- A lightweight triton-based General Matrix Multiplication (GEMM) library.☆67Jul 21, 2026Updated 3 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- HRX: Hip Runtime Extended☆20Updated this week
- High-performance GEMM kernel examples with FlyDSL on AMD GPUs.☆28Updated this week
- Automating analysis from trace files☆87Updated this week
- ☆102Nov 22, 2025Updated 8 months ago
- Generating Efficient AI-Centric Kernels☆148Updated this week
- ☆69Updated this week
- ☆30Updated this week
- ☆10May 15, 2024Updated 2 years ago
- Framework to reduce autotune overhead to zero for well known deployments.☆101Sep 19, 2025Updated 10 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Examples illustrating usage of the rocBLAS library☆17Aug 12, 2024Updated 2 years ago
- ☆16Jul 18, 2026Updated 3 weeks ago
- ☆19Mar 29, 2026Updated 4 months ago
- Public benchmark results from Kernel Arena, a leaderboard for LLM-generated AI accelerator kernels.☆21Mar 11, 2026Updated 5 months ago
- Repository to host ROCm Developer Hub Notebook Tutorials☆95Jun 29, 2026Updated last month
- ☆21Mar 17, 2026Updated 4 months ago
- Repository with examples and exercises for OLCF and AMD's HIP training series☆17Oct 16, 2023Updated 2 years ago
- Agents, and RL environment, for optimizing GPU kernels on AMD ROCm using LLM agents. Benchmarks LLM serving workloads end-to-end, profile…☆75Updated this week
- [DEPRECATED] Moved to ROCm/rocm-systems repo☆30May 28, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Optimizing diffusion for production-ready speeds☆40Jan 10, 2026Updated 7 months ago
- A GPU Cluster Simulator for Distributed Deep Learning Training.☆10Jan 15, 2022Updated 4 years ago
- ☆19Nov 11, 2025Updated 9 months ago
- [DEPRECATED] Moved to ROCm/rocm-systems repo☆145Updated this week
- AI Tensor Engine for ROCm☆524Updated this week
- ☆20May 30, 2026Updated 2 months ago
- ☆19May 9, 2025Updated last year
- ☆19Jun 6, 2025Updated last year
- NVFP4 Flash-Attention 4 on BlackWell☆38Jul 23, 2026Updated 2 weeks ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Wave: Python Domain-Specific Language for High Performance Machine Learning☆57Jun 29, 2026Updated last month
- Evaluating Large Language Models for CUDA Code Generation ComputeEval is a framework designed to generate and evaluate CUDA code from Lar…☆144May 19, 2026Updated 2 months ago
- ☆353Updated this week
- LLVM/MLIR based compiler instrumentation of AMD GPU kernels☆21Jul 13, 2025Updated last year
- FLA but cuTile☆27Apr 17, 2026Updated 3 months ago
- Ring network model test to demonstrate the use of CoreNEURON☆11Jul 5, 2026Updated last month
- Parallel Tasking Library (PTL) - Lightweight C++11 mutilthreading tasking system featuring thread-pool, task-groups, and lock-free task q…☆47Nov 14, 2024Updated last year