Open-source framework for the research and development of foundation models.
☆1,215Jul 21, 2026Updated this week
Alternatives and similar repositories for marin
Users that are interested in marin are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Legible, Scalable, Reproducible Foundation Models with Named Tensors and Jax☆707Jan 26, 2026Updated 5 months ago
- NanoGPT (124M) in 90 seconds☆5,548Jul 3, 2026Updated 2 weeks ago
- PyTorch building blocks for the OLMo ecosystem☆1,410Updated this week
- SkyRL: A Modular Full-stack RL Library for LLMs☆2,085Updated this week
- Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.☆1,761Updated this week
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Minimal yet performant LLM examples in pure JAX☆267Jul 4, 2026Updated 2 weeks ago
- slime is an LLM post-training framework for RL Scaling.☆7,569Updated this week
- Agentic RL Training at Scale☆1,702Updated this week
- Official repository for Parallax (Parameterized Local Linear Attention)☆65Jul 7, 2026Updated 2 weeks ago
- MoE training for Me and You and maybe other people☆394Mar 15, 2026Updated 4 months ago
- The simplest, fastest repository for training/finetuning medium-sized GPTs.☆199Jan 19, 2026Updated 6 months ago
- 100M tokens. Infinite compute. Lowest val loss wins.☆514Jul 3, 2026Updated 2 weeks ago
- Physics of Language Models: Part 4.2, Canon Layers at Scale where Synthetic Pretraining Resonates in Reality☆356May 20, 2026Updated 2 months ago
- 🚀 Efficient implementations for emerging model architectures☆5,388Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Simple & Scalable Pretraining for Neural Architecture Research☆337Mar 31, 2026Updated 3 months ago
- Accelerating MoE with IO and Tile-aware Optimizations☆732Jul 4, 2026Updated 2 weeks ago
- Minimalistic large language model 3D-parallelism training☆2,760May 26, 2026Updated last month
- ☆70Apr 8, 2026Updated 3 months ago
- Our library for RL environments + evals☆4,390Updated this week
- [NeurIPS 2025 Spotlight] Reasoning Environments for Reinforcement Learning with Verifiable Rewards☆1,463Apr 17, 2026Updated 3 months ago
- Tokamax: A GPU and TPU kernel library.☆249Updated this week
- Tile primitives for speedy kernels☆3,555Jul 13, 2026Updated last week
- Home for "How To Scale Your Model", a short blog-style textbook about scaling LLMs on TPUs☆1,285Jul 13, 2026Updated last week
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Dion optimizer algorithm☆496Jul 12, 2026Updated last week
- Named Tensors for Legible Deep Learning in JAX☆226Nov 8, 2025Updated 8 months ago
- Muon is an optimizer for hidden layers in neural networks☆2,721May 24, 2026Updated last month
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework☆22,587Updated this week
- Post-training with Tinker☆3,869Updated this week
- Efficient optimizers☆335Jul 11, 2026Updated last week
- 🧱 Modula software package☆337Aug 18, 2025Updated 11 months ago
- Minimalistic 4D-parallelism distributed training framework for education purpose☆2,255Aug 26, 2025Updated 10 months ago
- AllenAI's post-training codebase☆3,803Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for "What really matters in matrix-whitening optimizers?"☆25Oct 31, 2025Updated 8 months ago
- Measuring how well CLI agents like Claude Code or Codex CLI can post-train base LLMs on a single H100 GPU in 10 hours☆463Updated this week
- A Lightweight LLM Post-Training Library☆2,378Updated this week
- A framework for few-shot evaluation of language models.☆13,359Jul 13, 2026Updated last week
- Muon is Scalable for LLM Training☆1,508Aug 3, 2025Updated 11 months ago
- A Quirky Assortment of CuTe Kernels☆1,064Updated this week
- Minimal but scalable implementation of large language models in JAX☆34Nov 28, 2025Updated 7 months ago