A list of awesome compiler projects and papers for tensor computation and deep learning.
☆2,770Oct 19, 2024Updated last year
Alternatives and similar repositories for awesome-tensor-compilers
Users that are interested in awesome-tensor-compilers are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Must read research papers and links to tools and datasets that are related to using machine learning for compilers and systems optimisati…☆1,685Jan 21, 2026Updated 6 months ago
- compiler learning resources collect.☆2,766May 20, 2026Updated 2 months ago
- A flexible and efficient deep neural network (DNN) compiler that generates high-performance executable from a DNN model description.☆1,002Sep 19, 2024Updated last year
- A list of tutorials, paper, talks, and open-source projects for emerging compiler and architecture☆535Jan 15, 2025Updated last year
- Open Machine Learning Compiler Framework☆13,665Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- BladeDISC is an end-to-end DynamIc Shape Compiler project for machine learning workloads.☆932Dec 30, 2024Updated last year
- An MLIR-based compiler framework bridges DSLs (domain-specific languages) to DSAs (domain-specific architectures).☆750Updated this week
- Dive into Deep Learning Compiler☆651Jun 19, 2022Updated 4 years ago
- A retargetable MLIR-based machine learning compiler and runtime toolkit.☆3,889Updated this week
- The Torch-MLIR project aims to provide first class support from the PyTorch ecosystem to the MLIR ecosystem.☆1,885Updated this week
- ☆193Mar 28, 2023Updated 3 years ago
- The Tensor Algebra SuperOptimizer for Deep Learning☆744Jan 26, 2023Updated 3 years ago
- ☆2,033Jul 29, 2023Updated 3 years ago
- Automatic Mapping Generation, Verification, and Exploration for ISA-based Spatial Accelerators☆125Oct 26, 2022Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Development repository for the Triton language and compiler☆19,942Updated this week
- [MLSys 2021] IOS: Inter-Operator Scheduler for CNN Acceleration☆201Apr 27, 2022Updated 4 years ago
- Distributed Compiler based on Triton for Parallel Systems☆1,518Updated this week
- CUDA Templates and Python DSLs for High-Performance Linear Algebra☆10,250Updated this week
- SparseTIR: Sparse Tensor Compiler for Deep Learning☆145Mar 31, 2023Updated 3 years ago
- 🚀 Awesome System for Machine Learning ⚡️ AI System Papers and Industry Practice. ⚡️ System for Machine Learning, LLM (Large Language Mod…☆4,274Jul 25, 2025Updated last year
- Automatic Schedule Exploration and Optimization Framework for Tensor Computations☆184Apr 25, 2022Updated 4 years ago
- An open-source efficient deep learning framework/compiler, written in python.☆744Sep 4, 2025Updated 11 months ago
- Hands-On Practical MLIR Tutorial☆822Oct 20, 2023Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆101Nov 4, 2022Updated 3 years ago
- A model compilation solution for various hardware☆476Aug 20, 2025Updated 11 months ago
- MegCC是一个运行时超轻量,高效,移植简单的深度学习模型编译器☆483Oct 23, 2024Updated last year
- A Easy-to-understand TensorOp Matmul Tutorial☆450Mar 5, 2026Updated 5 months ago
- FlashInfer: Kernel Library for LLM Serving☆6,159Updated this week
- ☆421Feb 24, 2026Updated 5 months ago
- A home for the final text of all TVM RFCs.☆111Sep 24, 2024Updated last year
- row-major matmul optimization☆750May 14, 2026Updated 3 months ago
- Mirage Persistent Kernel: Compiling LLMs into a MegaKernel☆2,423Aug 4, 2026Updated last week
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Representation and Reference Lowering of ONNX Models in MLIR Compiler Infrastructure☆1,040Updated this week
- ☆637Apr 5, 2026Updated 4 months ago
- PET: Optimizing Tensor Programs with Partially Equivalent Transformations and Automated Corrections☆126Jun 23, 2022Updated 4 years ago
- how to optimize some algorithm in cuda.☆3,200Updated this week
- ☆249Jul 27, 2025Updated last year
- MLIR For Beginners tutorial☆1,342Jul 18, 2025Updated last year
- Awesome resources for GPUs☆639Mar 10, 2026Updated 5 months ago