The Art of Debugging Open Book
☆1,677Jul 31, 2026Updated this week
Alternatives and similar repositories for the-art-of-debugging
Users that are interested in the-art-of-debugging are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Machine Learning Engineering Open Book☆18,509Updated this week
- Stas' Python Cookbook - Python recipes that I use daily☆57Jul 2, 2026Updated last month
- GPU programming related news and material links☆2,253Jun 15, 2026Updated last month
- ML/DL Math and Method notes☆66Dec 2, 2023Updated 2 years ago
- Material for gpu-mode lectures☆6,388Jun 15, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Minimalistic large language model 3D-parallelism training☆2,771May 26, 2026Updated 2 months ago
- Python tools☆14Oct 22, 2023Updated 2 years ago
- Solve puzzles. Improve your pytorch.☆4,261Jul 15, 2024Updated 2 years ago
- What would you do with 1000 H100s...☆1,185Jan 10, 2024Updated 2 years ago
- Tile primitives for speedy kernels☆3,588Jul 13, 2026Updated 3 weeks ago
- Deep learning for dummies. All the practical details and useful utilities that go into working with real models.☆845Mar 15, 2026Updated 4 months ago
- A subset of PyTorch's neural network modules, written in Python using OpenAI's Triton.☆604May 13, 2026Updated 2 months ago
- An ML Systems Onboarding list☆1,110Feb 19, 2026Updated 5 months ago
- Puzzles for learning Triton☆2,549Apr 1, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Minimalistic 4D-parallelism distributed training framework for education purpose☆2,266Aug 26, 2025Updated 11 months ago
- Solve puzzles. Learn CUDA.☆12,373Sep 1, 2024Updated last year
- Pragmatic approach to parsing import profiles for CI's☆12Jul 1, 2024Updated 2 years ago
- A PyTorch native platform for training generative AI models☆5,581Updated this week
- A compact implementation of SGLang, designed to demystify the complexities of modern LLM serving systems.☆4,668May 17, 2026Updated 2 months ago
- Simple and efficient pytorch-native transformer text generation in <1000 LOC of python.☆6,235Aug 22, 2025Updated 11 months ago
- PyTorch Single Controller☆1,066Updated this week
- Code, labs, and resources for O'Reilly AI Systems Performance Engineering: GPU optimization, distributed training, inference scaling, and…☆1,763Jul 6, 2026Updated 3 weeks ago
- ☆353Updated this week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Sharing both practical insights and theoretical knowledge about LLM evaluation that we gathered while managing the Open LLM Leaderboard a…☆2,133Dec 3, 2025Updated 8 months ago
- MoE training for Me and You and maybe other people☆396Mar 15, 2026Updated 4 months ago
- Quantized LLM training in pure CUDA/C++.☆252Jul 23, 2026Updated last week
- ☆591Jul 11, 2024Updated 2 years ago
- A playbook for systematically maximizing the performance of deep learning models.☆30,270Jun 18, 2024Updated 2 years ago
- A curriculum for learning about gpu performance engineering, from scratch to what the frontier AI labs do☆1,278Apr 27, 2026Updated 3 months ago
- Development repository for the Triton language and compiler☆19,844Updated this week
- My solutions for Advanced Python Mastery (course by @dabeaz)☆11Jan 29, 2024Updated 2 years ago
- SGLang is a high-performance serving framework for large language models and multimodal models.☆31,163Updated this week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Nano vLLM☆14,790Apr 26, 2026Updated 3 months ago
- Minimal example scripts of the Hugging Face Trainer, focused on staying under 150 lines☆197May 6, 2024Updated 2 years ago
- FlashInfer: Kernel Library for LLM Serving☆6,090Updated this week
- Puffing up reinforcement learning☆6,225Updated this week
- A pure-Python implementation of the Nvidia CuTe layout algebra intended to be approachable and easy to learn.☆236Jun 29, 2026Updated last month
- My learning notes for ML SYS.☆6,810Updated this week
- Efficient Triton Kernels for LLM Training☆6,543Updated this week