Open-source CUDA, Triton and HIP compiler targeting multiple GPU and CPU architectures.
☆1,749Sep 14, 2026Updated last week
Alternatives and similar repositories for Booth
Users that are interested in Booth are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Monte Carlo neutron transport in C99. GPU via Booth (AMD MI300X, NVIDIA RTX). ENDF/B-VII.1 nuclear data. Validated against ICSBEP benchma…☆15Aug 21, 2026Updated last month
- CUDA on non-NVIDIA GPUs☆14,869Sep 2, 2026Updated 2 weeks ago
- An MLIR-based compiler that takes GPU kernels and compiles them to real hardware instructions. Interactive web visualizer included.☆145Mar 21, 2026Updated 6 months ago
- ☆1,088May 18, 2025Updated last year
- CUDA Tile IR is an MLIR-based intermediate representation and compiler infrastructure for CUDA kernel optimization, focusing on tile-base…☆1,034Sep 10, 2026Updated last week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- High-efficiency LLM inference engine in C++/CUDA. Run Llama 70B on RTX 3090.☆466Feb 22, 2026Updated 6 months ago
- nCPU: model-native and tensor-optimized CPU research runtimes with organized workloads, tools, and docs☆658Jul 30, 2026Updated last month
- cuda-oxide is a Rust-to-CUDA compiler that lets you write (SIMT) GPU kernels in safe(ish), idiomatic Rust. It compiles standard Rust code…☆3,538Updated this week
- Fast and Furious AMD Kernels☆470Updated this week
- Minimal x86 Kernel - built in Zig☆230Feb 19, 2026Updated 7 months ago
- Super fast FP32 matrix multiplication on RDNA3☆92Mar 30, 2025Updated last year
- Tracing JIT Deep Learning Library☆31Nov 6, 2025Updated 10 months ago
- RDNA3 emulator☆63Apr 16, 2026Updated 5 months ago
- Use your NVIDIA GPU's VRAM as swap space on Linux. Built for laptops with soldered memory and no upgrade path. If you have an RTX card si…☆564Jul 17, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A novel take on lossless data compression☆3,183Updated this week
- build-once run-anywhere c library☆21,298Jul 20, 2026Updated 2 months ago
- Tensor library & inference framework for machine learning☆118Oct 3, 2025Updated 11 months ago
- cuTile Rust provides a safe, tile-based kernel programming DSL for the Rust programming language. It features a safe host-side API for pa…☆1,005Updated this week
- AI Tensor Engine for ROCm☆565Updated this week
- Tensor library for machine learning☆15,383Sep 14, 2026Updated last week
- Multi-platform high-performance compute language extension for Rust.☆2,385Updated this week
- A Linux framework to enable userspace-defined "Virtual" PCIe card shims to enable in-host PCIe card driver development.☆374Sep 9, 2026Updated last week
- Full-throttle, wire-speed hardware implementation of Wireguard VPN, using low-cost Artix7 FPGA with opensource toolchain. If you seek sec…☆1,362Sep 14, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Performance focused header-only container library. Currently primarily contains a fast B+Tree implementation.☆75Jan 6, 2026Updated 8 months ago
- Fil-C: completely compatible memory safety for C and C++☆3,845Updated this week
- A very fast linker for Linux☆3,987Updated this week
- A debugger for Linux☆1,723Jul 9, 2026Updated 2 months ago
- Voxtral ASR & TTS running natively and in the browser. A Rust implementation of Mistral's Voxtral mini realtime ASR / TTS using the Burn …☆822Apr 2, 2026Updated 5 months ago
- A detour through the Linux dynamic linker☆556Jul 13, 2025Updated last year
- You like pytorch? You like micrograd? You love tinygrad! ❤️☆33,633Updated this week
- A minimal GPU design in Verilog to learn how GPUs work from the ground up☆13,003Aug 18, 2024Updated 2 years ago
- Pure C inference of Mistral Voxtral Realtime 4B speech to text model☆1,750Feb 15, 2026Updated 7 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Flux 2 image generation model pure C inference☆1,992Feb 13, 2026Updated 7 months ago
- Ecosystem of libraries and tools for writing and executing fast GPU code fully in Rust.☆5,388Updated this week
- tiniest x86-64-linux emulator☆7,587Dec 10, 2025Updated 9 months ago
- Any model. Any hardware. Zero compromise. Built with @ziglang / @openxla / MLIR / @bazelbuild☆4,072Updated this week
- Build your own high performance LLM inference engine in C++ and CUDA - a smaller version of vLLM☆1,126Updated this week
- chipStar is a tool for compiling and running HIP/CUDA on SPIR-V via OpenCL or Level Zero APIs.☆378Updated this week
- LLM inference in C/C++☆129,071Updated this week