Ring-V2 is a reasoning MoE LLM provided and open-sourced by InclusionAI.
☆99Oct 23, 2025Updated 11 months ago
Alternatives and similar repositories for Ring-V2
Users that are interested in Ring-V2 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Ring is a reasoning MoE LLM provided and open-sourced by InclusionAI, derived from Ling.☆108Aug 5, 2025Updated last year
- Ling-V2 is a MoE LLM provided and open-sourced by InclusionAI.☆281Oct 4, 2025Updated 11 months ago
- A high-performance kernel library for LLM training☆89Apr 28, 2026Updated 4 months ago
- Revealing and unlocking the context boundary of reward models☆21May 10, 2026Updated 4 months ago
- ☆47Feb 28, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆15Feb 26, 2026Updated 6 months ago
- Code for the paper "Stack Attention: Improving the Ability of Transformers to Model Hierarchical Patterns"☆19Mar 15, 2024Updated 2 years ago
- ☆25Aug 20, 2025Updated last year
- A comprehensive and efficient long-context model evaluation framework☆31Aug 31, 2026Updated 3 weeks ago
- [ICLR 2026] Geometric-Mean Policy Optimization☆105Jan 26, 2026Updated 7 months ago
- ☆28Feb 28, 2026Updated 6 months ago
- Defeating the Training-Inference Mismatch via FP16☆197Nov 14, 2025Updated 10 months ago
- BigCodeArena: Unveiling More Reliable Human Preferences in Code Generation via Execution☆62Oct 13, 2025Updated 11 months ago
- context denoising training for long-context modeling☆17Oct 10, 2025Updated 11 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Code for the preprint "Cache Me If You Can: How Many KVs Do You Need for Effective Long-Context LMs?"☆48Jul 29, 2025Updated last year
- ☆74Oct 23, 2025Updated 11 months ago
- Code and data to explore neural scaling laws of xLSTM and Transformer models.☆24Apr 8, 2026Updated 5 months ago
- Long Context Research☆38Aug 31, 2026Updated 3 weeks ago
- Evaluating the faithfulness of long-context language models☆30Oct 21, 2024Updated last year
- Analysis code for Neurips 2025 paper "SciArena: An Open Evaluation Platform for Foundation Models in Scientific Literature Tasks"☆57Aug 6, 2025Updated last year
- PaCoRe: Learning to Scale Test-Time Compute with Parallel Coordinated Reasoning☆341Feb 5, 2026Updated 7 months ago
- PresentAgent-2: Towards Generalist Multimodal Presentation Agents☆18Jun 5, 2026Updated 3 months ago
- CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning☆37Aug 28, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Official repository of paper "RNNs Are Not Transformers (Yet): The Key Bottleneck on In-context Retrieval"☆27Apr 17, 2024Updated 2 years ago
- Code release for Ming-UniVision: Joint Image Understanding and Geneation with a Continuous Unified Tokenizer☆143Oct 14, 2025Updated 11 months ago
- The officalimplement of dLLM-Factory☆25Jul 12, 2025Updated last year
- A high-performance RL training-inference weight synchronization framework, designed to enable second-level parameter updates from trainin…☆177Updated this week
- Fluid Language Model Benchmarking☆29Sep 16, 2025Updated last year
- Landing repository for the paper "Predicting the Order of Upcoming Tokens Improves Language Modeling"☆48May 13, 2026Updated 4 months ago
- ☆14Jun 13, 2025Updated last year
- [ICLR 2025] Official Pytorch Implementation of "Mix-LN: Unleashing the Power of Deeper Layers by Combining Pre-LN and Post-LN" by Pengxia…☆30Jul 24, 2025Updated last year
- This repositories contains the reference implementation for the Sparse Delta Memory paper.More precisely, it contains the model definitio…☆39Jul 9, 2026Updated 2 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Ling is a MoE LLM provided and open-sourced by InclusionAI.☆267May 14, 2025Updated last year
- ☆16Aug 5, 2025Updated last year
- ☆561May 25, 2026Updated 3 months ago
- The official repo for "OpenMoE 2: Sparse Diffusion Language Models".☆58Dec 28, 2025Updated 8 months ago
- [ICML 2025] Predictive Data Selection: The Data That Predicts Is the Data That Teaches☆66Mar 4, 2025Updated last year
- ☆12Sep 16, 2025Updated last year
- Implementation of Fast Weight Attention☆35Updated this week