A high-performance RL training-inference weight synchronization framework, designed to enable second-level parameter updates from training to inference in RL workflows
☆169Aug 14, 2026Updated this week
Alternatives and similar repositories for Awex
Users that are interested in Awex are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆42Jul 23, 2026Updated 3 weeks ago
- A NCCL extension library, designed to efficiently offload GPU memory allocated by the NCCL communication library.☆116Dec 17, 2025Updated 8 months ago
- An asynchronous streaming data management module for efficient post-training.☆133Updated this week
- APRIL: Active Partial Rollouts in Reinforcement Learning to Tame Long-tail Generation. A system-level optimization for scalable LLM tra…☆60Oct 11, 2025Updated 10 months ago
- Checkpoint-engine is a simple middleware to update model weights in LLM inference engines☆998Updated this week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [Archived] For the latest updates and community contribution, please visit: https://github.com/Ascend/TransferQueue or https://gitcode.co…☆15Updated this week
- An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale☆571Updated this week
- Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.☆2,014Updated this week
- An industrial extension library of pytorch to accelerate large scale model training☆62Aug 13, 2025Updated last year
- Standardized environment infrastructure for Agentic AI development.☆314Jul 10, 2026Updated last month
- ☆46Sep 8, 2025Updated 11 months ago
- [ASPLOS'26] Taming the Long-Tail: Efficient Reasoning RL Training with Adaptive Drafter☆175Feb 27, 2026Updated 5 months ago
- Composable and Embeddable Communication Runtime for Distributed AI Services☆101Jun 5, 2026Updated 2 months ago
- Bridge Megatron-Core to Hugging Face/Reinforcement Learning☆229Jun 15, 2026Updated 2 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [NSDI'26] PolyRL is a reinforcement learning framework for LLM that harvest spot instances on the cloud to reduce cost.☆19Mar 30, 2026Updated 4 months ago
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.☆5,673Updated this week
- NVSHMEM‑Tutorial: Build a DeepEP‑like GPU Buffer☆198Feb 11, 2026Updated 6 months ago
- ☆382Jan 28, 2026Updated 6 months ago
- Code repository for the SOSP'25 paper DCP: Addressing Input Dynamism In Long-Context Training via Dynamic Context Parallelism.☆22Nov 28, 2025Updated 8 months ago
- ☆46Oct 15, 2025Updated 10 months ago
- SiMM: Scalable in-Memory Middleware☆42Apr 20, 2026Updated 3 months ago
- Awesome system papers for AI☆21Updated this week
- FA4-based Relative Attention Kernel developed by TML and Colfax☆18Jul 17, 2026Updated last month
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Allow torch tensor memory to be released and resumed later☆270Updated this week
- slime is an LLM post-training framework for RL Scaling.☆8,093Updated this week
- Provide performance insight capabilities for RL frameworks.☆57Aug 11, 2026Updated last week
- verl Zero-Mismatch Dense/MoE HuggingFace Rollout☆64Aug 5, 2026Updated last week
- Compact and Agent-Native MoE Training System☆333Jul 31, 2026Updated 2 weeks ago
- On demand communication☆33Apr 16, 2026Updated 4 months ago
- Accelerating MoE with IO and Tile-aware Optimizations☆743Updated this week
- A fast communication-overlapping library for tensor/expert parallelism on GPUs.☆1,352Aug 28, 2025Updated 11 months ago
- A lightweight, AI-native training framework for large language models. Designed for fast iteration, reproducible experiments, and modular…☆584May 18, 2026Updated 3 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆914Updated this week
- NexRL is an ultra-loosely-coupled LLM post-training framework.☆118Jul 30, 2026Updated 2 weeks ago
- Training library for Megatron-based models with bidirectional Hugging Face conversion capability☆863Updated this week
- Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.☆6,300Updated this week
- Ring-V2 is a reasoning MoE LLM provided and open-sourced by InclusionAI.☆99Oct 23, 2025Updated 9 months ago
- Perplexity GPU Kernels☆600Nov 7, 2025Updated 9 months ago
- ☆1,305May 20, 2026Updated 2 months ago