Long-RL: Scaling RL to Long Sequences (NeurIPS 2025)
β726Sep 24, 2025Updated 10 months ago
Alternatives and similar repositories for Long-RL
Users that are interested in Long-RL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR 2026]QeRL enables RL for 32B LLMs on a single H100 GPU.β511Mar 30, 2026Updated 3 months ago
- Video-R1: Reinforcing Video Reasoning in MLLMs [π₯the first paper to explore R1 for video]β882Dec 14, 2025Updated 7 months ago
- Long Video Gen Infrastructureβ2,491Jul 15, 2026Updated last week
- β1,250Nov 20, 2025Updated 8 months ago
- Official Code for "Mini-o3: Scaling Up Reasoning Patterns and Interaction Turns for Visual Search"β422Jan 29, 2026Updated 5 months ago
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- EasyR1: An Efficient, Scalable, Multi-Modality RL Training Framework based on veRLβ5,081Updated this week
- VeOmni: Scaling Any Modality Model Training with Model-Centric Distributed Recipe Zooβ2,107Updated this week
- MMaDA - Open-Sourced Multimodal Large Diffusion Language Models (dLLMs with block diffusion, mixed-CoT, unified RL)β1,660Feb 14, 2026Updated 5 months ago
- Official implementation of BLIP3o-Seriesβ1,664Nov 29, 2025Updated 7 months ago
- Open-source unified multimodal modelβ6,116May 4, 2026Updated 2 months ago
- [NeurIPS 2025] An official implementation of Flow-GRPO: Training Flow Matching Models via Online RLβ2,430May 7, 2026Updated 2 months ago
- Native Multimodal Models are World Learnersβ1,537Dec 30, 2025Updated 6 months ago
- [ICLR & NeurIPS 2025] Repository for Show-o series, One Single Transformer to Unify Multimodal Understanding and Generation.β1,964Jan 8, 2026Updated 6 months ago
- Seed1.5-VL, a vision-language foundation model designed to advance general-purpose multimodal understanding and reasoning, achieving statβ¦β1,583Jun 14, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- A fork to add multimodal model training to open-r1β1,593Feb 8, 2025Updated last year
- Structured Video Comprehension of Real-World Shortsβ239Sep 21, 2025Updated 10 months ago
- [TPAMI 2026] Ego-R1: Agentic Chain-of-Tool-Thought for Ultra-Long Egocentric Video Reasoningβ165Jun 10, 2026Updated last month
- Official implementation of UnifiedReward & [NeurIPS 2025] UnifiedReward-Think & UnifiedReward-Flexβ796Jun 18, 2026Updated last month
- [NeurIPS 2025] Efficient Reasoning Vision Language Modelsβ460Sep 18, 2025Updated 10 months ago
- [NeurIPS2025] The official implementation of MindOmni: Unleashing Reasoning Generation in Vision Language Models with RGPOβ139Oct 15, 2025Updated 9 months ago
- Kimi-VL: Mixture-of-Experts Vision-Language Model for Multimodal Reasoning, Long-Context Understanding, and Strong Agent Capabilitiesβ1,206Jul 15, 2025Updated last year
- Fully Open Framework for Democratized Multimodal Trainingβ1,149Updated this week
- The official code of "Thinking With Videos: Multimodal Tool-Augmented Reinforcement Learning for Long Video Reasoning"β102Oct 15, 2025Updated 9 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- β¨First Open-Source R1-like Video-LLM [2025/02/18]β382Jul 1, 2026Updated 3 weeks ago
- Next-Token Prediction is All You Needβ2,433Jan 12, 2026Updated 6 months ago
- A version of verl to support diverse tool use [TMLR 2026]β1,024Jul 15, 2026Updated last week
- Official repo and evaluation implementation of VSI-Benchβ734Aug 5, 2025Updated 11 months ago
- [CVPR 2026] LongVT: Incentivizing "Thinking with Long Videos" via Native Tool Callingβ255Jun 24, 2026Updated last month
- Official implementation of "Fast-dLLM: Training-free Acceleration of Diffusion LLM by Enabling KV Cache and Parallel Decoding"β1,063May 30, 2026Updated last month
- VILA is a family of state-of-the-art vision language models (VLMs) for diverse multimodal AI tasks across the edge, data center, and clouβ¦β3,845Mar 12, 2026Updated 4 months ago
- β4,713Jun 15, 2026Updated last month
- Long Context Transfer from Language to Visionβ407Mar 18, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Cambrian-S: Towards Spatial Supersensing in Videoβ562Apr 3, 2026Updated 3 months ago
- [ACL-2026] MMSearch-R1 is an end-to-end RL framework that enables LMMs to perform on-demand, multi-turn search with real-world multimodalβ¦β470Apr 7, 2026Updated 3 months ago
- [CVPR 2026] Video-as-Answer: Predict and Generate Next Video Event with Joint-GRPOβ119Feb 28, 2026Updated 4 months ago
- [ICCV 2025] GameFactory: Creating New Games with Generative Interactive Videosβ489Mar 22, 2025Updated last year
- Official Implementation of Paper Transfer between Modalities with MetaQueriesβ325Oct 12, 2025Updated 9 months ago
- Cambrian-1 is a family of multimodal LLMs with a vision-centric design.β2,008Nov 7, 2025Updated 8 months ago
- MiMo-VLβ641Aug 21, 2025Updated 11 months ago