The Art of Debugging Open Book
☆1,721Sep 3, 2026Updated last week
Alternatives and similar repositories for the-art-of-debugging
Users that are interested in the-art-of-debugging are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Machine Learning Engineering Open Book☆18,979Updated this week
- Stas' Python Cookbook - Python recipes that I use daily☆65Jul 2, 2026Updated 2 months ago
- GPU programming related news and material links☆2,329Jun 15, 2026Updated 2 months ago
- ML/DL Math and Method notes☆67Dec 2, 2023Updated 2 years ago
- Material for gpu-mode lectures☆6,581Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Minimalistic large language model 3D-parallelism training☆2,820May 26, 2026Updated 3 months ago
- Python tools☆14Oct 22, 2023Updated 2 years ago
- Solve puzzles. Improve your pytorch.☆4,323Jul 15, 2024Updated 2 years ago
- What would you do with 1000 H100s...☆1,196Jan 10, 2024Updated 2 years ago
- Tile primitives for speedy kernels☆3,675Updated this week
- Deep learning for dummies. All the practical details and useful utilities that go into working with real models.☆852Aug 12, 2026Updated last month
- A subset of PyTorch's neural network modules, written in Python using OpenAI's Triton.☆604Aug 14, 2026Updated 3 weeks ago
- An ML Systems Onboarding list☆1,124Feb 19, 2026Updated 6 months ago
- Puzzles for learning Triton☆2,592Apr 1, 2026Updated 5 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Minimalistic 4D-parallelism distributed training framework for education purpose☆2,299Aug 26, 2025Updated last year
- Solve puzzles. Learn CUDA.☆12,455Sep 1, 2024Updated 2 years ago
- Pragmatic approach to parsing import profiles for CI's☆12Jul 1, 2024Updated 2 years ago
- A PyTorch native platform for training generative AI models☆5,728Updated this week
- A compact implementation of SGLang, designed to demystify the complexities of modern LLM serving systems.☆5,032May 17, 2026Updated 3 months ago
- Simple and efficient pytorch-native transformer text generation in <1000 LOC of python.☆6,251Aug 22, 2025Updated last year
- PyTorch Single Controller☆1,076Updated this week
- Code, labs, and resources for O'Reilly AI Systems Performance Engineering: GPU optimization, distributed training, inference scaling, and…☆1,952Updated this week
- ☆355Updated this week
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Sharing both practical insights and theoretical knowledge about LLM evaluation that we gathered while managing the Open LLM Leaderboard a…☆2,145Dec 3, 2025Updated 9 months ago
- A curated resource list for learning AI performance engineering, from GPU fundamentals to production inference.☆2,741Aug 23, 2026Updated 3 weeks ago
- MoE training for Me and You and maybe other people☆399Mar 15, 2026Updated 5 months ago
- Quantized LLM training in pure CUDA/C++.☆258Aug 31, 2026Updated last week
- ☆592Jul 11, 2024Updated 2 years ago
- Development repository for the Triton language and compiler☆20,139Updated this week
- A playbook for systematically maximizing the performance of deep learning models.☆30,319Jun 18, 2024Updated 2 years ago
- My solutions for Advanced Python Mastery (course by @dabeaz)☆11Jan 29, 2024Updated 2 years ago
- SGLang is a high-performance serving framework for large language models and multimodal models.☆35,850Updated this week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Nano vLLM☆15,411Apr 26, 2026Updated 4 months ago
- Minimal example scripts of the Hugging Face Trainer, focused on staying under 150 lines☆196May 6, 2024Updated 2 years ago
- FlashInfer: Kernel Library for LLM Serving☆6,384Updated this week
- Puffing up reinforcement learning☆6,342Updated this week
- A pure-Python implementation of the Nvidia CuTe layout algebra intended to be approachable and easy to learn.☆242Jun 29, 2026Updated 2 months ago
- My learning notes for ML SYS.☆7,331Updated this week
- Efficient Triton Kernels for LLM Training☆6,610Updated this week