Minimal and scalable research codebase in JAX, designed for rapid iteration on frontier research in LLM and other autoregressive models.
☆556Aug 6, 2026Updated last week
Alternatives and similar repositories for simply
Users that are interested in simply are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MoE training for Me and You and maybe other people☆396Mar 15, 2026Updated 5 months ago
- ☆304Jul 15, 2024Updated 2 years ago
- A transformer that executes a one-instruction Turing-complete computer — two approaches: hand-coded weights (no training) and learned fro…☆41Mar 3, 2026Updated 5 months ago
- Minimal yet performant LLM examples in pure JAX☆273Jul 4, 2026Updated last month
- Our library for RL environments + evals☆4,521Updated this week
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Agentic RL Training at Scale☆1,943Updated this week
- ☆27Jan 22, 2026Updated 6 months ago
- AI-Driven Scientific, Algorithmic, and Systems Discovery☆608Updated this week
- autonomous nanogpt optimizer speedrun☆110May 14, 2026Updated 3 months ago
- The Automated LLM Speedrunning Benchmark measures how well LLM agents can reproduce previous innovations and discover new ones in languag…☆146May 6, 2026Updated 3 months ago
- Puffing up reinforcement learning☆6,277Updated this week
- Open-source framework for the research and development of foundation models.☆1,267Updated this week
- NanoGPT (124M) in 90 seconds☆5,681Aug 9, 2026Updated last week
- Compile programs directly into transformer weights. Includes a 2D convex-hull KV cache with O(log n) inference.☆215Jun 1, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- WIP☆96Aug 13, 2024Updated 2 years ago
- Dion optimizer algorithm☆535Updated this week
- Deriving steepest descent convergence bounds and hyperparameter scaling laws in machine learning optimization from first principles, form…☆16Apr 11, 2026Updated 4 months ago
- Pure C / AVX-512 port of Craftax-Classic. 47.8M SPS on a Ryzen 9 9950X3D -- 3.2x an RTX Pro 6000 Blackwell on the same env.☆24Jul 14, 2026Updated last month
- ☆21Jul 7, 2026Updated last month
- slime is an LLM post-training framework for RL Scaling.☆8,119Updated this week
- A Python DSL to write Nvidia PTX for Hopper and Blackwell in JAX and PyTorch☆374Jul 9, 2026Updated last month
- Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.☆2,028Updated this week
- Lightweight Zotero MCP server for AI agents☆15May 16, 2026Updated 3 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆31Mar 13, 2026Updated 5 months ago
- A simple, performant and scalable JAX-based world modeling codebase.☆162Jan 15, 2026Updated 7 months ago
- The best ChatGPT that $100 can buy.☆58Updated this week
- Simple & Scalable Pretraining for Neural Architecture Research☆343Mar 31, 2026Updated 4 months ago
- Kernels, of the mega variety :)☆804May 26, 2026Updated 2 months ago
- Lab Cookbook☆42Aug 5, 2026Updated 2 weeks ago
- Checkpoint-engine is a simple middleware to update model weights in LLM inference engines☆999Aug 12, 2026Updated last week
- 100M tokens. Infinite compute. Lowest val loss wins.☆526Jul 3, 2026Updated last month
- A Gym for Agentic LLMs☆506Jan 21, 2026Updated 6 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [NeurIPS 2025 Spotlight] Reasoning Environments for Reinforcement Learning with Verifiable Rewards☆1,487Apr 17, 2026Updated 4 months ago
- A Quirky Assortment of CuTe Kernels☆1,117Updated this week
- Tokamax: A GPU and TPU kernel library.☆265Updated this week
- Post-training with Tinker☆4,031Updated this week
- SkyRL: A Modular Full-stack RL Library for LLMs☆2,167Updated this week
- Tile primitives for speedy kernels☆3,633Jul 13, 2026Updated last month
- Physics of Language Models: Part 4.2, Canon Layers at Scale where Synthetic Pretraining Resonates in Reality☆357May 20, 2026Updated 2 months ago