Wave: Python Domain-Specific Language for High Performance Machine Learning
β58Jun 29, 2026Updated 3 weeks ago
Alternatives and similar repositories for wave
Users that are interested in wave are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- HRX: Hip Runtime Extendedβ18Updated this week
- ASTER π« : Assembly Tooling and Representationsβ32Jul 1, 2026Updated 2 weeks ago
- A lightweight triton-based General Matrix Multiplication (GEMM) library.β65Jun 13, 2026Updated last month
- β19Jun 6, 2025Updated last year
- Fuzz testing for Dafnyβ13Jul 7, 2022Updated 4 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits β’ AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- FlyDSL is the Python frontβend of the project: Flexible LaYout DSL.β237Updated this week
- AMD RAD's multi-GPU Triton-based framework for seamless multi-GPU programmingβ193Updated this week
- A collection of out-of-tree extensions for the Triton language and compilerβ30Updated this week
- A Triton-only attention backend for vLLMβ27Updated this week
- Ship correct and fast LLM kernels to PyTorchβ151Jan 14, 2026Updated 6 months ago
- β32Jul 2, 2025Updated last year
- Super fast FP32 matrix multiplication on RDNA3β92Mar 30, 2025Updated last year
- Interactive version of the CuTe layout paperβ57Apr 14, 2026Updated 3 months ago
- Ahead of Time (AOT) Triton Math Libraryβ100Jul 13, 2026Updated last week
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- FLA but cuTileβ27Apr 17, 2026Updated 3 months ago
- incubator repo for CUDA-TileIR backendβ148Jul 10, 2026Updated last week
- IREE's PyTorch Frontend, based on Torch Dynamo.β109Jul 1, 2026Updated 2 weeks ago
- MLIR-based partitioning systemβ198Updated this week
- β61Updated this week
- [DEPRECATED] Moved to ROCm/rocm-libraries repoβ140Jul 13, 2026Updated last week
- β183Updated this week
- β54Updated this week
- TPP experimentation on MLIR for linear algebraβ155Updated this week
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A pure-Python implementation of the Nvidia CuTe layout algebra intended to be approachable and easy to learn.β231Jun 29, 2026Updated 3 weeks ago
- β21Mar 17, 2026Updated 4 months ago
- CUDA Tile IR is an MLIR-based intermediate representation and compiler infrastructure for CUDA kernel optimization, focusing on tile-baseβ¦β999Jul 6, 2026Updated 2 weeks ago
- Exercises for Learning MLIR (Originally written for PPoPP 2026)β106Feb 5, 2026Updated 5 months ago
- C++ Graph API and JIT Engine powered by IREEβ25Jul 1, 2026Updated 2 weeks ago
- Surgical GPU kernel benchmark: 7 hard problems, frontier coding agents, roofline-graded against hardware peak.β19Jun 12, 2026Updated last month
- Code for "An Introduction to Tensor Tiling in MLIR" tutorial given at EuroLLVM 2025β24Jun 5, 2025Updated last year
- AiTer Optimized Modelβ141Updated this week
- Embedded Universal DSL: a good DSL for us, by usβ76Updated this week
- GPUs on demand by Runpod - Special Offer Available β’ AdRun AI, ML, and HPC workloads on powerful cloud GPUsβwithout limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Github mirror of trition-lang/triton repo.β178Updated this week
- A Triton JIT runtime and ffi provider in C++β37Updated this week
- Code snippets and reproductions from JustAByteβ48Apr 6, 2026Updated 3 months ago
- Library to interface Compilers and ML models for ML-Enabled Compiler Optimizationsβ20Oct 19, 2025Updated 9 months ago
- β30Jun 16, 2026Updated last month
- Languages, Tools, and Techniques for Accelerator Designβ33Nov 2, 2021Updated 4 years ago
- Public benchmark results from Kernel Arena, a leaderboard for LLM-generated AI accelerator kernels.β20Mar 11, 2026Updated 4 months ago