Development repository for the Triton language and compiler
☆146Sep 22, 2026Updated this week
Alternatives and similar repositories for triton
Users that are interested in triton are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Ahead of Time (AOT) Triton Math Library☆100Updated this week
- Fast and memory-efficient exact attention☆239Aug 12, 2026Updated last month
- [DEPRECATED] Moved to ROCm/rocm-libraries repo☆114Updated this week
- [DEPRECATED] Moved to ROCm/rocm-libraries repo. NOTE: develop branch is maintained as a read-only mirror☆548Updated this week
- [DEPRECATED] Moved to ROCm/rocm-libraries repo☆142Sep 15, 2026Updated last week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆76Updated this week
- ☆18Apr 10, 2026Updated 5 months ago
- AI Tensor Engine for ROCm☆567Updated this week
- ☆192Sep 15, 2026Updated last week
- This is the AMD-maintained fork of the LLVM git repository. This repository accepts pull requests and issues related to AMD fork-specific…☆235Updated this week
- [DEPRECATED] Moved to ROCm/rocm-libraries repo☆27Sep 16, 2026Updated last week
- A high-throughput and memory-efficient inference and serving engine for LLMs☆124Updated this week
- LLVM/MLIR based compiler instrumentation of AMD GPU kernels☆21Jul 13, 2025Updated last year
- CMake modules used within the ROCm libraries☆78Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- AITemplate is a Python framework which renders neural network into high performance CUDA/HIP C++ code. Specialized for FP16 TensorCore (N…☆12Jun 24, 2024Updated 2 years ago
- A tool for generating information about the matrix multiplication instructions in AMD Radeon™ and AMD Instinct™ accelerators☆145Apr 10, 2026Updated 5 months ago
- python package of rocm-smi-lib☆25Dec 15, 2025Updated 9 months ago
- ☆75Updated this week
- FlyDSL is the Python front‑end of the project: a Flexible Layout Python DSL for expressing tiling, partitioning, data movement, and kerne…☆284Updated this week
- AMD RAD's multi-GPU Triton-based framework for seamless multi-GPU programming☆202Sep 17, 2026Updated last week
- [DEPRECATED] Moved to ROCm/rocm-systems repo☆153May 28, 2026Updated 3 months ago
- [DEPRECATED] Moved to ROCm/rocm-systems repo☆167May 28, 2026Updated 3 months ago
- ☆20Oct 11, 2023Updated 2 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- vLLM: A high-throughput and memory-efficient inference and serving engine for LLMs☆94Updated this week
- 8-bit CUDA functions for PyTorch☆71Aug 19, 2026Updated last month
- [DEPRECATED] Moved to ROCm/rocm-systems repo☆83May 28, 2026Updated 3 months ago
- A PyTorch Extension: Tools for easy mixed precision and distributed training in Pytorch☆24Aug 19, 2026Updated last month
- Hackable and optimized Transformers building blocks, supporting a composable construction.☆34May 29, 2026Updated 3 months ago
- Super fast FP32 matrix multiplication on RDNA3☆92Mar 30, 2025Updated last year
- [DEPRECATED] Moved to ROCm/rocm-libraries repo☆262Sep 10, 2026Updated 2 weeks ago
- FlagGems is an operator library for large language models implemented in the Triton Language.☆1,125Updated this week
- Composable transformations of Python+NumPy programs: differentiate, vectorize, JIT to GPU/TPU, and more☆32Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- AMD's graph optimization engine.☆334Updated this week
- Inference Llama 2 with a model compiled to native code by TorchInductor☆14Feb 8, 2024Updated 2 years ago
- [DEPRECATED] Moved to ROCm/rocm-libraries repo☆402Updated this week
- extensible collectives library in triton☆97Mar 31, 2025Updated last year
- Framework to reduce autotune overhead to zero for well known deployments.☆101Sep 19, 2025Updated last year
- [DEPRECATED] Moved to ROCm/rocm-libraries repo☆153Sep 15, 2026Updated last week
- [DEPRECATED] Moved to ROCm/rocm-systems repo☆278Sep 16, 2026Updated last week