Ling-V2 is a MoE LLM provided and open-sourced by InclusionAI.
☆279Oct 4, 2025Updated 11 months ago
Alternatives and similar repositories for Ling-V2
Users that are interested in Ling-V2 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Ling is a MoE LLM provided and open-sourced by InclusionAI.☆267May 14, 2025Updated last year
- Ring-V2 is a reasoning MoE LLM provided and open-sourced by InclusionAI.☆99Oct 23, 2025Updated 10 months ago
- A high-performance kernel library for LLM training☆88Apr 28, 2026Updated 4 months ago
- Ring is a reasoning MoE LLM provided and open-sourced by InclusionAI, derived from Ling.☆108Aug 5, 2025Updated last year
- ☆46Feb 28, 2026Updated 6 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Mixture-of-Basis-Experts for Compressing MoE-based LLMs☆38Dec 24, 2025Updated 8 months ago
- Muon is Scalable for LLM Training☆1,542Aug 3, 2025Updated last year
- ☆34Aug 14, 2026Updated 3 weeks ago
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.☆5,723Updated this week
- CUDA kernels for linear attention variants, written in CuTe DSL and CUTLASS C++.☆544Updated this week
- Best practices for training DeepSeek, Mixtral, Qwen and other MoE models using Megatron Core.☆202May 29, 2026Updated 3 months ago
- dUltra: Ultra-Fast Diffusion Large Language Models via Reinforcement Learning☆17Jul 11, 2026Updated last month
- ☆1,602Nov 17, 2025Updated 9 months ago
- The official implemention of "Depth-Breadth Synergy in RLVR: Unlocking LLM Reasoning Gains with Adaptive Exploration" (ICML 2026)☆24Feb 4, 2026Updated 7 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A Recipe for Building LLM Reasoners to Solve Complex Instructions☆32Oct 9, 2025Updated 10 months ago
- VeOmni: Scaling Any Modality Model Training with Model-Centric Distributed Recipe Zoo☆2,192Updated this week
- ABench is an evolving open-source benchmark suite designed to rigorously evaluate and enhance Large Language Models (LLMs) on complex cro…☆29Jul 30, 2026Updated last month
- A high-performance RL training-inference weight synchronization framework, designed to enable second-level parameter updates from trainin…☆174Aug 20, 2026Updated 2 weeks ago
- The open-source code of MetaStone-S1.☆106Aug 1, 2025Updated last year
- The official repo for “Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem” [EMNLP25]☆33Sep 1, 2025Updated last year
- Structured Primitives for Efficient Architecture Research☆21Dec 22, 2025Updated 8 months ago
- Landing repository for the paper "Predicting the Order of Upcoming Tokens Improves Language Modeling"☆48May 13, 2026Updated 3 months ago
- LLM Inference with Microscaling Format☆35Nov 12, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆271Aug 28, 2026Updated last week
- Accelerating MoE with IO and Tile-aware Optimizations☆757Aug 29, 2026Updated last week
- Don't just regulate gradients like in Muon, regulate the weights too☆32Jul 30, 2025Updated last year
- The official implementation of dLLM-Var☆35Nov 6, 2025Updated 9 months ago
- ☆27Feb 28, 2026Updated 6 months ago
- Revisiting Mid-training in the Era of Reinforcement Learning Scaling☆188Jul 23, 2025Updated last year
- dInfer: An Efficient Inference Framework for Diffusion Language Models☆479Feb 11, 2026Updated 6 months ago
- PaCoRe: Learning to Scale Test-Time Compute with Parallel Coordinated Reasoning☆339Feb 5, 2026Updated 7 months ago
- ☆48Sep 8, 2025Updated 11 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆1,644Nov 18, 2025Updated 9 months ago
- Search, understand, reproduce, and improve an idea with ease☆1,229Updated this week
- slime is an LLM post-training framework for RL Scaling.☆8,381Updated this week
- ☆131Sep 9, 2025Updated 11 months ago
- Overflow Prevention Enhances Long-Context Recurrent LLMs (COLM 2025)☆18Jul 8, 2025Updated last year
- [WWW 2026 Oral] MoE-CL:Self-Evolving LLMs via Continual Instruction Tuning☆22Dec 1, 2025Updated 9 months ago
- Easy and Efficient dLLM Fine-Tuning☆268Aug 4, 2026Updated last month