Minimalistic 4D-parallelism distributed training framework for education purpose
☆2,315Aug 26, 2025Updated last year
Alternatives and similar repositories for picotron
Users that are interested in picotron are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Minimalistic large language model 3D-parallelism training☆2,828Sep 23, 2026Updated last week
- ☆262Nov 24, 2025Updated 10 months ago
- A PyTorch native platform for training generative AI models☆5,777Updated this week
- The simplest, fastest repository for training/finetuning small-sized VLMs.☆5,040Oct 27, 2025Updated 11 months ago
- A compact implementation of SGLang, designed to demystify the complexities of modern LLM serving systems.☆5,210May 17, 2026Updated 4 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Efficient Triton Kernels for LLM Training☆6,644Updated this week
- NanoGPT (124M) in 90 seconds☆5,900Updated this week
- Nano vLLM☆15,694Apr 26, 2026Updated 5 months ago
- Tile primitives for speedy kernels☆3,737Sep 12, 2026Updated 3 weeks ago
- FlexAttention based, minimal vllm-style inference engine for fast Gemma 2 inference.☆361Nov 2, 2025Updated 11 months ago
- slime is an LLM post-training framework for RL Scaling.☆8,589Updated this week
- Implementing DeepSeek R1's GRPO algorithm from scratch☆1,904Apr 18, 2025Updated last year
- 🚀 Efficient implementations for emerging model architectures☆5,811Updated this week
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework☆23,738Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Our library for RL environments + evals☆4,670Updated this week
- My learning notes for ML SYS.☆7,424Sep 20, 2026Updated last week
- FlashInfer: Kernel Library for LLM Serving☆6,538Updated this week
- Simple MPI implementation for prototyping or learning☆330Aug 6, 2025Updated last year
- Puzzles for learning Triton☆2,621Apr 1, 2026Updated 6 months ago
- What would you do with 1000 H100s...☆1,196Jan 10, 2024Updated 2 years ago
- Distributed Compiler and Optimized Parallel Kernels☆1,555Sep 18, 2026Updated 2 weeks ago
- Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.☆3,040Updated this week
- Meta Lingua: a lean, efficient, and easy-to-hack codebase to research LLMs.☆4,768Jul 18, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- PyTorch native quantization for training and inference☆2,991Updated this week
- Ongoing research training transformer models at scale☆18,061Updated this week
- Material for gpu-mode lectures☆6,677Sep 9, 2026Updated 3 weeks ago
- SGLang is a high-performance serving framework for large language models and multimodal models.☆36,748Updated this week
- Agentic RL Training at Scale☆2,117Updated this week
- Helpful tools and examples for working with flex-attention☆1,253Updated this week
- Train transformer language models with reinforcement learning.☆19,443Updated this week
- ☆1,315May 20, 2026Updated 4 months ago
- Simple and efficient pytorch-native transformer text generation in <1000 LOC of python.☆6,259Aug 22, 2025Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on H…☆3,564Updated this week
- Everything about the SmolLM and SmolVLM family of models☆3,915Sep 23, 2026Updated last week
- Ring attention implementation with flash attention☆1,061Sep 10, 2025Updated last year
- A scalable asynchronous reinforcement learning implementation with in-flight weight updates.☆433Aug 5, 2026Updated last month
- GPU programming related news and material links☆2,355Jun 15, 2026Updated 3 months ago
- 🚀 Efficiently (pre)training foundation models with native PyTorch features, including FSDP for training and SDPA implementation of Flash…☆288Nov 24, 2025Updated 10 months ago
- A subset of PyTorch's neural network modules, written in Python using OpenAI's Triton.☆604Aug 14, 2026Updated last month