Minimalistic 4D-parallelism distributed training framework for education purpose
☆2,274Aug 26, 2025Updated 11 months ago
Alternatives and similar repositories for picotron
Users that are interested in picotron are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Minimalistic large language model 3D-parallelism training☆2,779May 26, 2026Updated 2 months ago
- ☆255Nov 24, 2025Updated 8 months ago
- A PyTorch native platform for training generative AI models☆5,603Updated this week
- The simplest, fastest repository for training/finetuning small-sized VLMs.☆4,979Oct 27, 2025Updated 9 months ago
- A compact implementation of SGLang, designed to demystify the complexities of modern LLM serving systems.☆4,710May 17, 2026Updated 2 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Efficient Triton Kernels for LLM Training☆6,556Updated this week
- NanoGPT (124M) in 90 seconds☆5,648Aug 2, 2026Updated last week
- Nano vLLM☆14,910Apr 26, 2026Updated 3 months ago
- Tile primitives for speedy kernels☆3,616Jul 13, 2026Updated 3 weeks ago
- FlexAttention based, minimal vllm-style inference engine for fast Gemma 2 inference.☆358Nov 2, 2025Updated 9 months ago
- slime is an LLM post-training framework for RL Scaling.☆7,813Updated this week
- Implementing DeepSeek R1's GRPO algorithm from scratch☆1,895Apr 18, 2025Updated last year
- 🚀 Efficient implementations for emerging model architectures☆5,524Updated this week
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework☆22,869Updated this week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Our library for RL environments + evals☆4,475Updated this week
- My learning notes for ML SYS.☆6,838Updated this week
- FlashInfer: Kernel Library for LLM Serving☆6,129Updated this week
- Simple MPI implementation for prototyping or learning☆327Aug 6, 2025Updated last year
- Puzzles for learning Triton☆2,554Apr 1, 2026Updated 4 months ago
- What would you do with 1000 H100s...☆1,188Jan 10, 2024Updated 2 years ago
- Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.☆1,927Updated this week
- Distributed Compiler based on Triton for Parallel Systems☆1,511Updated this week
- Meta Lingua: a lean, efficient, and easy-to-hack codebase to research LLMs.☆4,763Jul 18, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- PyTorch native quantization and sparsity for training and inference☆2,932Updated this week
- Ongoing research training transformer models at scale☆17,367Updated this week
- Material for gpu-mode lectures☆6,408Jun 15, 2026Updated last month
- SGLang is a high-performance serving framework for large language models and multimodal models.☆31,535Updated this week
- Agentic RL Training at Scale☆1,855Updated this week
- Helpful tools and examples for working with flex-attention☆1,221Updated this week
- Train transformer language models with reinforcement learning.☆19,027Updated this week
- ☆1,302May 20, 2026Updated 2 months ago
- Simple and efficient pytorch-native transformer text generation in <1000 LOC of python.☆6,240Aug 22, 2025Updated 11 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on H…☆3,483Updated this week
- Everything about the SmolLM and SmolVLM family of models☆3,865May 26, 2026Updated 2 months ago
- Ring attention implementation with flash attention☆1,045Sep 10, 2025Updated 10 months ago
- A scalable asynchronous reinforcement learning implementation with in-flight weight updates.☆432Updated this week
- GPU programming related news and material links☆2,260Jun 15, 2026Updated last month
- Fast and memory-efficient exact attention☆24,654Updated this week
- 🚀 Efficiently (pre)training foundation models with native PyTorch features, including FSDP for training and SDPA implementation of Flash…☆288Nov 24, 2025Updated 8 months ago