Train LLM from scratch for $5 USD - Research.
☆211Feb 23, 2026Updated 5 months ago
Alternatives and similar repositories for 5-dollar-llm
Users that are interested in 5-dollar-llm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆86Dec 13, 2025Updated 7 months ago
- ☆16Jun 15, 2026Updated last month
- ☆30Dec 15, 2025Updated 7 months ago
- ☆27Jul 19, 2026Updated 3 weeks ago
- Cuda implemenation of flash-kmeans, 2x faster☆23May 8, 2026Updated 3 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆15Jul 20, 2026Updated 3 weeks ago
- NanoGPT (124M) as fast as possible☆20Apr 15, 2025Updated last year
- Push open-source AI research to the frontier by the end of 2026 - open frontier science for everybody.☆29Aug 13, 2025Updated 11 months ago
- Structured Primitives for Efficient Architecture Research☆21Dec 22, 2025Updated 7 months ago
- Streamline on-policy/off-policy distillation workflows in a few lines of code☆109Updated this week
- ☆27Jun 7, 2026Updated 2 months ago
- Repository of implementations of classic and sota rl algorithms from scratch in PyTorch☆226Aug 3, 2026Updated last week
- Like ARC, but code to generate visual puzzles. 1D puzzles first.☆23Aug 17, 2024Updated last year
- A truly open version of gpt-oss which shows the entire pre-training from scratch☆91Sep 4, 2025Updated 11 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆39Aug 4, 2025Updated last year
- Various test models in WNNX format. It can view with `pip install wnetron && wnetron`☆12Jun 22, 2022Updated 4 years ago
- Meta-Reinforcement Learning with Self-Reflection☆33Mar 26, 2026Updated 4 months ago
- Source code for the paper "Positional Attention: Expressivity and Learnability of Algorithmic Computation"☆14May 26, 2025Updated last year
- CS194-196 Course Project☆14Feb 20, 2025Updated last year
- UM1 test programs and sample code☆11Jul 25, 2022Updated 4 years ago
- A repository consisting of paper/architecture replications of classic/SOTA AI/ML papers in pytorch☆424Nov 11, 2025Updated 8 months ago
- documenting my work in inference engineering☆25Apr 19, 2026Updated 3 months ago
- GPT for FACodec☆13Mar 25, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- MoE training for Me and You and maybe other people☆396Mar 15, 2026Updated 4 months ago
- A simple uv workspace☆19Apr 5, 2025Updated last year
- Gecko Architecture☆18Jan 13, 2026Updated 6 months ago
- Development of New Neural Network Optimizers☆15Nov 27, 2025Updated 8 months ago
- A curated list of awesome resources, libraries, frameworks, and tools for multi-agent systems (MAS) research and development.☆33Feb 17, 2025Updated last year
- Qwen3-0.6B megakernel: 527 tok/s decode on RTX 3090 (3.8x faster than PyTorch)☆131Feb 10, 2026Updated 6 months ago
- ☆17May 6, 2025Updated last year
- ☆19Mar 3, 2025Updated last year
- An educational distributed training and inference library for neural nets using local computing☆74Jun 10, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Companion code for The Physics of LLM Inference book☆26Apr 21, 2026Updated 3 months ago
- 🥪 Mess portal where owners can set their weekly menu, price, time, and students can purchase their desired coupons, with a QR code syste…☆11Jun 2, 2023Updated 3 years ago
- ☆21Feb 10, 2025Updated last year
- Marketplace ML experiment - training without backprop☆28Sep 9, 2025Updated 11 months ago
- j1-micro (1.7B) & j1-nano (600M) are absurdly tiny but mighty reward models.☆106Jul 19, 2025Updated last year
- Data pipeline for HRM-Text pretraining☆70May 21, 2026Updated 2 months ago
- ☆69Mar 21, 2025Updated last year