Open-source CUDA, Triton and HIP compiler targeting multiple GPU and CPU architectures.
☆1,717Jul 17, 2026Updated this week
Alternatives and similar repositories for Booth
Users that are interested in Booth are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Monte Carlo neutron transport in C99. GPU via Booth (AMD MI300X, NVIDIA RTX). ENDF/B-VII.1 nuclear data. Validated against ICSBEP benchma…☆15Jun 5, 2026Updated last month
- CUDA on non-NVIDIA GPUs☆14,624Updated this week
- An MLIR-based compiler that takes GPU kernels and compiles them to real hardware instructions. Interactive web visualizer included.☆139Mar 21, 2026Updated 4 months ago
- ☆1,087May 18, 2025Updated last year
- High-efficiency LLM inference engine in C++/CUDA. Run Llama 70B on RTX 3090.☆465Feb 22, 2026Updated 4 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- CUDA Tile IR is an MLIR-based intermediate representation and compiler infrastructure for CUDA kernel optimization, focusing on tile-base…☆999Jul 6, 2026Updated 2 weeks ago
- nCPU: model-native and tensor-optimized CPU research runtimes with organized workloads, tools, and docs☆652Jul 11, 2026Updated last week
- cuda-oxide is an experimental Rust-to-CUDA compiler that lets you write (SIMT) GPU kernels in safe(ish), idiomatic Rust. It compiles stan…☆2,940Updated this week
- Fast and Furious AMD Kernels☆444Jul 10, 2026Updated last week
- Minimal x86 Kernel - built in Zig☆228Feb 19, 2026Updated 5 months ago
- Super fast FP32 matrix multiplication on RDNA3☆92Mar 30, 2025Updated last year
- Tracing JIT Deep Learning Library☆29Nov 6, 2025Updated 8 months ago
- Use your NVIDIA GPU's VRAM as swap space on Linux. Built for laptops with soldered memory and no upgrade path. If you have an RTX card si…☆508Updated this week
- A novel data compression framework☆3,139Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Tensor library & inference framework for machine learning☆118Oct 3, 2025Updated 9 months ago
- build-once run-anywhere c library☆21,170Updated this week
- AI Tensor Engine for ROCm☆497Updated this week
- cuTile Rust provides a safe, tile-based kernel programming DSL for the Rust programming language. It features a safe host-side API for pa…☆706Updated this week
- Multi-platform high-performance compute language extension for Rust.☆2,278Updated this week
- A Linux framework to enable userspace-defined "Virtual" PCIe card shims to enable in-host PCIe card driver development.☆362Jul 14, 2026Updated last week
- Full-throttle, wire-speed hardware implementation of Wireguard VPN, using low-cost Artix7 FPGA with opensource toolchain. If you seek sec…☆1,333Updated this week
- Performance focused header-only container library. Currently primarily contains a fast B+Tree implementation.☆74Jan 6, 2026Updated 6 months ago
- Fil-C: completely compatible memory safety for C and C++☆3,433Updated this week
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Voxtral ASR & TTS running natively and in the browser. A Rust implementation of Mistral's Voxtral mini realtime ASR / TTS using the Burn …☆809Apr 2, 2026Updated 3 months ago
- A debugger for Linux☆1,698Jul 9, 2026Updated last week
- A very fast linker for Linux☆3,771Updated this week
- A detour through the Linux dynamic linker☆542Jul 13, 2025Updated last year
- Pure C inference of Mistral Voxtral Realtime 4B speech to text model☆1,710Feb 15, 2026Updated 5 months ago
- You like pytorch? You like micrograd? You love tinygrad! ❤️☆33,306Updated this week
- Flux 2 image generation model pure C inference☆1,967Feb 13, 2026Updated 5 months ago
- tiniest x86-64-linux emulator☆7,557Dec 10, 2025Updated 7 months ago
- LLM inference in C/C++☆121,053Updated this week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Ecosystem of libraries and tools for writing and executing fast GPU code fully in Rust.☆5,274Apr 29, 2026Updated 2 months ago
- Public repository of the Micro QuickJS Javascript Engine☆6,038Jun 4, 2026Updated last month
- chipStar is a tool for compiling and running HIP/CUDA on SPIR-V via OpenCL or Level Zero APIs.☆364Updated this week
- A machine learning accelerator core designed for energy-efficient AI at the edge.☆2,467Updated this week
- Build your own high performance LLM inference engine in C++ and CUDA - a smaller version of vLLM☆938Jul 2, 2026Updated 2 weeks ago
- A curated list of best cuda programming books☆940May 19, 2026Updated 2 months ago
- CUDA Templates and Python DSLs for High-Performance Linear Algebra☆10,104Updated this week