A high-performance RL training-inference weight synchronization framework, designed to enable second-level parameter updates from training to inference in RL workflows
☆168Jul 23, 2026Updated this week
Alternatives and similar repositories for Awex
Users that are interested in Awex are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆42Updated this week
- A NCCL extension library, designed to efficiently offload GPU memory allocated by the NCCL communication library.☆113Dec 17, 2025Updated 7 months ago
- An asynchronous streaming data management module for efficient post-training.☆123Jul 12, 2026Updated 2 weeks ago
- APRIL: Active Partial Rollouts in Reinforcement Learning to Tame Long-tail Generation. A system-level optimization for scalable LLM tra…☆60Oct 11, 2025Updated 9 months ago
- Checkpoint-engine is a simple middleware to update model weights in LLM inference engines☆984Jul 4, 2026Updated 3 weeks ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [Archived] For the latest updates and community contribution, please visit: https://github.com/Ascend/TransferQueue or https://gitcode.co…☆16Jan 16, 2026Updated 6 months ago
- An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale☆546Updated this week
- Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.☆1,805Updated this week
- An industrial extension library of pytorch to accelerate large scale model training☆62Aug 13, 2025Updated 11 months ago
- Standardized environment infrastructure for Agentic AI development.☆313Jul 10, 2026Updated 2 weeks ago
- ☆47Sep 8, 2025Updated 10 months ago
- [ASPLOS'26] Taming the Long-Tail: Efficient Reasoning RL Training with Adaptive Drafter☆174Feb 27, 2026Updated 5 months ago
- Composable and Embeddable Communication Runtime for Distributed AI Services☆102Jun 5, 2026Updated last month
- Bridge Megatron-Core to Hugging Face/Reinforcement Learning☆228Jun 15, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [NSDI'26] PolyRL is a reinforcement learning framework for LLM that harvest spot instances on the cloud to reduce cost.☆19Mar 30, 2026Updated 3 months ago
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.☆5,612Updated this week
- NVSHMEM‑Tutorial: Build a DeepEP‑like GPU Buffer☆195Feb 11, 2026Updated 5 months ago
- ☆380Jan 28, 2026Updated 6 months ago
- Code repository for the SOSP'25 paper DCP: Addressing Input Dynamism In Long-Context Training via Dynamic Context Parallelism.☆21Nov 28, 2025Updated 8 months ago
- ☆45Oct 15, 2025Updated 9 months ago
- SiMM: Scalable in-Memory Middleware☆41Apr 20, 2026Updated 3 months ago
- Awesome system papers for AI☆21Updated this week
- FA4-based Relative Attention Kernel developed by TML and Colfax☆17Jul 17, 2026Updated last week
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Allow torch tensor memory to be released and resumed later☆261Updated this week
- slime is an LLM post-training framework for RL Scaling.☆7,679Updated this week
- Provide performance insight capabilities for RL frameworks.☆48Updated this week
- verl Zero-Mismatch Dense/MoE HuggingFace Rollout☆63Jul 15, 2026Updated last week
- Compact and Agent-Native MoE Training System☆304Updated this week
- On demand communication☆34Apr 16, 2026Updated 3 months ago
- Accelerating MoE with IO and Tile-aware Optimizations☆732Jul 4, 2026Updated 3 weeks ago
- A fast communication-overlapping library for tensor/expert parallelism on GPUs.☆1,348Aug 28, 2025Updated 11 months ago
- ☆716Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A lightweight, AI-native training framework for large language models. Designed for fast iteration, reproducible experiments, and modular…☆579May 18, 2026Updated 2 months ago
- NexRL is an ultra-loosely-coupled LLM post-training framework.☆114Updated this week
- A comprehensive AI & ML project portfolio from the University of Texas at Austin PG Program, demonstrating real-world data science and ma…☆17Jan 25, 2026Updated 6 months ago
- Training library for Megatron-based models with bidirectional Hugging Face conversion capability☆833Updated this week
- Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.☆6,063Updated this week
- Ring-V2 is a reasoning MoE LLM provided and open-sourced by InclusionAI.☆98Oct 23, 2025Updated 9 months ago
- Perplexity GPU Kernels☆595Nov 7, 2025Updated 8 months ago