An experimental CPU backend for Triton (https//github.com/openai/triton)
☆48Aug 18, 2025Updated last year
Alternatives and similar repositories for triton-cpu
Users that are interested in triton-cpu are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- OpenAI Triton backend for Intel® GPUs☆268Updated this week
- Collection of scripts to build PyTorch and the domain libraries from source.☆14Jul 9, 2026Updated last month
- ☆21Mar 3, 2025Updated last year
- Benchmarking PyTorch 2.0 different models☆20Mar 19, 2023Updated 3 years ago
- Automatic differentiation for Triton Kernels☆29Aug 12, 2025Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- AI-ML-NLP Task Group☆13Aug 10, 2023Updated 3 years ago
- Repository for AI model benchmarking on TT-Buda☆15Feb 9, 2026Updated 6 months ago
- A fork of tvm/unity☆14Aug 12, 2023Updated 3 years ago
- Shared Middle-Layer for Triton Compilation☆346Dec 5, 2025Updated 9 months ago
- TPP experimentation on MLIR for linear algebra☆162Updated this week
- FP4 MAC Array☆20Apr 14, 2024Updated 2 years ago
- RESPECT: Reinforcement Learning based Edge Scheduling on Pipelined Coral Edge TPUs (DAC'23)☆11Apr 13, 2023Updated 3 years ago
- tenstorrent kernel from twitch☆29Mar 16, 2024Updated 2 years ago
- RISC-V kernel step-by-step implmenetation☆12Aug 24, 2026Updated 2 weeks ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆16Jul 3, 2025Updated last year
- SMT-LIB benchmarks for shape computations from deep learning models in PyTorch☆18Dec 21, 2022Updated 3 years ago
- ☆36Updated this week
- Writing FLUX in Triton☆42Sep 22, 2024Updated last year
- An MLIR-based toy DL compiler for TVM Relay.☆62Oct 16, 2022Updated 3 years ago
- FlexAttention w/ FlashAttention3 Support☆27Oct 5, 2024Updated last year
- VOICEVOX COREで利用するonnxruntimeのビルドを行うリポジトリ☆16Jul 23, 2026Updated last month
- ☆65Apr 26, 2025Updated last year
- [NeurIPS 2023] Sparse Modular Activation for Efficient Sequence Modeling☆40Dec 2, 2023Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- This repository is related to Choukanzu WG of the Japan OSS Promotion Forum.☆18Jul 27, 2026Updated last month
- TritonParse: A Compiler Tracer, Visualizer, and Reproducer for Triton Kernels☆215Updated this week
- Intel® Extension for MLIR. A staging ground for MLIR dialects and tools for Intel devices using the MLIR toolchain.☆156Updated this week
- Inference Llama 2 with a model compiled to native code by TorchInductor☆14Feb 8, 2024Updated 2 years ago
- Towards a million-node RISC-V cluster.☆14Mar 6, 2025Updated last year
- 🔀 yet another mixture of experts☆23Jun 5, 2026Updated 3 months ago
- Backward compatible ML compute opset inspired by HLO/MHLO☆695Updated this week
- ☆10Apr 27, 2026Updated 4 months ago
- libexecinfo for `execinfo.h` in musl systems☆20Nov 2, 2021Updated 4 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Provide Docker build sequences of PyTorch for various environments.☆16May 26, 2021Updated 5 years ago
- Artifacts of EVT ASPLOS'24☆29Mar 6, 2024Updated 2 years ago
- ☆20Jun 4, 2024Updated 2 years ago
- Performance Prediction Toolkit for GPUs☆41Mar 21, 2022Updated 4 years ago
- A collection of GPU experiments and benchmarks for my personal understanding and research.☆42Aug 19, 2026Updated 2 weeks ago
- Typst template for IEEE Conference☆23Jul 22, 2024Updated 2 years ago
- Remote source nodes for NNStreamer pipelines without GStreamer dependencies☆17Updated this week