Minimalistic 4D-parallelism distributed training framework for education purpose
☆2,291Aug 26, 2025Updated last year
Alternatives and similar repositories for picotron
Users that are interested in picotron are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Minimalistic large language model 3D-parallelism training☆2,802May 26, 2026Updated 3 months ago
- ☆258Nov 24, 2025Updated 9 months ago
- A PyTorch native platform for training generative AI models☆5,674Updated this week
- The simplest, fastest repository for training/finetuning small-sized VLMs.☆5,000Oct 27, 2025Updated 10 months ago
- A compact implementation of SGLang, designed to demystify the complexities of modern LLM serving systems.☆4,896May 17, 2026Updated 3 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Efficient Triton Kernels for LLM Training☆6,593Updated this week
- NanoGPT (124M) in 90 seconds☆5,713Aug 9, 2026Updated 2 weeks ago
- Nano vLLM☆15,202Apr 26, 2026Updated 4 months ago
- Tile primitives for speedy kernels☆3,657Updated this week
- FlexAttention based, minimal vllm-style inference engine for fast Gemma 2 inference.☆359Nov 2, 2025Updated 9 months ago
- slime is an LLM post-training framework for RL Scaling.☆8,290Updated this week
- Implementing DeepSeek R1's GRPO algorithm from scratch☆1,897Apr 18, 2025Updated last year
- 🚀 Efficient implementations for emerging model architectures☆5,650Updated this week
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework☆23,177Updated this week
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Our library for RL environments + evals☆4,567Updated this week
- My learning notes for ML SYS.☆6,979Aug 19, 2026Updated last week
- FlashInfer: Kernel Library for LLM Serving☆6,273Updated this week
- Simple MPI implementation for prototyping or learning☆327Aug 6, 2025Updated last year
- Puzzles for learning Triton☆2,575Apr 1, 2026Updated 4 months ago
- What would you do with 1000 H100s...☆1,190Jan 10, 2024Updated 2 years ago
- Distributed Compiler and Optimized Parallel Kernels☆1,528Aug 12, 2026Updated 2 weeks ago
- Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.☆2,278Updated this week
- Meta Lingua: a lean, efficient, and easy-to-hack codebase to research LLMs.☆4,766Jul 18, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- PyTorch native quantization for training and inference☆2,959Updated this week
- Ongoing research training transformer models at scale☆17,652Updated this week
- Material for gpu-mode lectures☆6,512Jun 15, 2026Updated 2 months ago
- SGLang is a high-performance serving framework for large language models and multimodal models.☆32,623Updated this week
- Agentic RL Training at Scale☆1,984Updated this week
- Helpful tools and examples for working with flex-attention☆1,228Updated this week
- Train transformer language models with reinforcement learning.☆19,169Updated this week
- ☆1,307May 20, 2026Updated 3 months ago
- Simple and efficient pytorch-native transformer text generation in <1000 LOC of python.☆6,248Aug 22, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on H…☆3,505Updated this week
- Everything about the SmolLM and SmolVLM family of models☆3,885May 26, 2026Updated 3 months ago
- Ring attention implementation with flash attention☆1,049Sep 10, 2025Updated 11 months ago
- A scalable asynchronous reinforcement learning implementation with in-flight weight updates.☆433Aug 5, 2026Updated 3 weeks ago
- GPU programming related news and material links☆2,306Jun 15, 2026Updated 2 months ago
- Fast and memory-efficient exact attention☆24,794Updated this week
- 🚀 Efficiently (pre)training foundation models with native PyTorch features, including FSDP for training and SDPA implementation of Flash…☆288Nov 24, 2025Updated 9 months ago