Optimized GPU compiler for LLM inference. Choose from a list of optimized recipes or optimize your own model via kernel fusion, autotuning, and advanced scheduling. Run benchmarks across different GPU types and configurations, track results and share experiments with the community.
☆82Oct 2, 2026Updated this week
Alternatives and similar repositories for emmy
Users that are interested in emmy are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- World's first Nintendo 3DS emulator for Apple devices based on Citra.☆18Apr 7, 2023Updated 3 years ago
- Arithmetic multiplier benchmarks☆12Nov 13, 2017Updated 8 years ago
- Implementation of local search-based algorithms for solving SAT and Max-SAT in Python☆13Dec 6, 2020Updated 5 years ago
- A package dedicated for running benchmark agreement testing☆19Sep 18, 2025Updated last year
- A fork of the Kissat SAT solver with additional features. Supports incremental solving.☆17Aug 13, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Minimize server usage by leveraging a decentralized peer-to-peer network for ultra-low-latency live streaming among users.☆13Feb 19, 2024Updated 2 years ago
- Make FP4 on 5090 Great Again☆24Jul 20, 2026Updated 2 months ago
- A design automation framework to engineer decision diagrams yourself☆27Sep 9, 2026Updated 3 weeks ago
- axum_embed is a library that provides a service for serving embedded files using the axum web framework.☆20Jan 6, 2025Updated last year
- A modular Lustre to C / Horn clauses compiler☆22Nov 17, 2018Updated 7 years ago
- Benchmarking Intelligence Efficiency of LM Inference☆95Sep 9, 2026Updated 3 weeks ago
- Exercises for Learning MLIR (Originally written for PPoPP 2026)☆112Jul 21, 2026Updated 2 months ago
- My submission for the GPUMODE/AMD fp8 mm challenge☆29Jun 4, 2025Updated last year
- High-performance GPU kernels for LLM inference in OpenAI Triton. Fused RMSNorm, SwiGLU, INT8 GEMM with benchmarks and roofline analysis.☆45Jul 22, 2026Updated 2 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Lightweight Python Wrapper for OpenVINO, enabling LLM inference on NPUs☆30Dec 17, 2024Updated last year
- High-performance GPU kernels for Ads and Recsys model training, independently implemented and optimized for real-world workloads and mode…☆52Updated this week
- Collection of official scripts created by the Dione Team.☆15Feb 21, 2026Updated 7 months ago
- Volume Manipulation Library☆16Jul 13, 2023Updated 3 years ago
- Repository for the CUDA H100 Course☆76Apr 12, 2026Updated 5 months ago
- ☆13Sep 10, 2026Updated 3 weeks ago
- this is an easy way to make ai podcast useing ai loccaly like ollama and the tts of piper☆16Feb 17, 2026Updated 7 months ago
- Mistral Vibe rewritten in Rust by Devstral 2☆21Dec 23, 2025Updated 9 months ago
- Web frontend for Myria☆12Sep 30, 2020Updated 6 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A series of high-performance GEMM (General Matrix Multiply) implementations Iteratively optimised for H100 GPUs in Pure CUDA.☆83Feb 18, 2026Updated 7 months ago
- ☆11Jun 11, 2020Updated 6 years ago
- Official Problem Sets / Reference Kernels for the GPU MODE Leaderboard!☆310Sep 18, 2026Updated 2 weeks ago
- In-process, multi master, distributed database☆21Jun 18, 2026Updated 3 months ago
- GPEmu, a GPU emulator for faster and cheaper prototyping and evaluation of deep learning system research☆45Dec 2, 2024Updated last year
- ☆30Jan 26, 2023Updated 3 years ago
- Type inference implementation in OCaml using Algorithm W☆10Aug 26, 2021Updated 5 years ago
- Audio Video development Kit, supporting audio、video、IPC、door bell、speech recognition...☆20Jun 6, 2025Updated last year
- Back in a Minute☆11Jul 19, 2019Updated 7 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Local LLM Server Manager + LlaMA.cpp + Chat☆141Updated this week
- PIDX☆14Jan 20, 2020Updated 6 years ago
- ☆35Jan 12, 2014Updated 12 years ago
- ☆16Aug 1, 2024Updated 2 years ago
- Reverse engineering NVIDIA SASS instruction dictionary, kernel audits and pattern recognition across GPU architectures.☆341May 18, 2026Updated 4 months ago
- Learn CUDA with PyTorch☆490Sep 25, 2026Updated last week
- In-memory key-value store for testing applications against weak behaviors of a database.☆11Feb 4, 2021Updated 5 years ago