MiroRL is an MCP-first reinforcement learning framework for deep research agent.
☆246Aug 27, 2025Updated 10 months ago
Alternatives and similar repositories for MiroRL
Users that are interested in MiroRL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MiroTrain is an efficient and algorithm-first framework research agent.☆142Aug 27, 2025Updated 10 months ago
- MiroMind-M1 is a fully open-source series of reasoning language models built on Qwen-2.5, focused on advancing mathematical reasoning.☆279Aug 12, 2025Updated 11 months ago
- 🏆 Top-1 on 5+ benchmarks | Web UI | Supports MiroThinker, Claude, Kimi, OpenAI☆3,077Jul 6, 2026Updated 2 weeks ago
- MiroThinker is a deep research agent optimized for complex research and prediction tasks. Our latest models, MiroThinker-1.7, achieves 74…☆8,345Jul 6, 2026Updated 2 weeks ago
- MiroEval: A benchmark and evaluation framework for deep research agents — 100 tasks (70 text, 30 multimodal) assessed across synthesis qu…☆46Jul 6, 2026Updated 2 weeks ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ICLR 2026] End-to-End Reinforcement Learning for Multi-Turn Tool-Integrated Reasoning☆401Mar 30, 2026Updated 3 months ago
- ☆90Aug 16, 2025Updated 11 months ago
- A version of verl to support diverse tool use [TMLR 2026]☆1,021Updated this week
- Scaling Deep Research via Reinforcement Learning in Real-world Environments.☆782May 10, 2026Updated 2 months ago
- NexRL is an ultra-loosely-coupled LLM post-training framework.☆114May 13, 2026Updated 2 months ago
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.☆5,579Updated this week
- Implementation for FP8/INT8 Rollout for RL training without performence drop.☆306Nov 7, 2025Updated 8 months ago
- Scaling RL on advanced reasoning models☆691Oct 20, 2025Updated 9 months ago
- slime is an LLM post-training framework for RL Scaling.☆7,569Updated this week
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Democratizing Reinforcement Learning for LLMs