🔍 OpenSearch-VL provides a fully open recipe for training strong multimodal deep search agents through high-quality data curation, diverse visual/search tools, and fatal-aware agentic reinforcement learning.
☆287Jul 30, 2026Updated last month
Alternatives and similar repositories for OpenSearch-VL
Users that are interested in OpenSearch-VL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 🐧 Unify-Agent: An end-to-end unified multimodal agent for faithful, knowledge-grounded image generation.☆92May 2, 2026Updated 4 months ago
- [ICLR 2026]🌴 ARES is an open-source framework for adaptive multimodal reasoning, featuring a two-stage pipeline—Adaptive Cold-Start and …☆23Feb 3, 2026Updated 7 months ago
- The official code of "Towards Long-horizon Agentic Multimodal Search"☆30Apr 17, 2026Updated 5 months ago
- SearchAgent-Zero: A Scalable Multi-Turn Search Agent RL Framework☆165Aug 27, 2026Updated 3 weeks ago
- HyperEyes is a parallel multimodal search agent that fuses visual grounding and retrieval into a single atomic action, enabling concurren…☆76May 23, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICML 2026 & EMNLP 2026] Multimodal deep-research MLLM and benchmark. The first long-horizon multimodal deep-research MLLM, extending the…☆684Aug 8, 2026Updated last month
- Official Code for "Mini-o3: Scaling Up Reasoning Patterns and Interaction Turns for Visual Search"☆426Jan 29, 2026Updated 7 months ago
- Agent-RRM: Exploring Reasoning Reward Model for Agents☆73Mar 17, 2026Updated 6 months ago
- ☆71Feb 27, 2026Updated 6 months ago
- [ACL-2026] MMSearch-R1 is an end-to-end RL framework that enables LMMs to perform on-demand, multi-turn search with real-world multimodal…☆486Apr 7, 2026Updated 5 months ago
- UniRL is a Framework for Unified Multimodal Model Reinforcement Learning☆975Updated this week
- Vero: An Open RL Recipe for General Visual Reasoning☆148Aug 29, 2026Updated 3 weeks ago
- OpenSeeker: A search agent with open-source data and models☆777Aug 26, 2026Updated 3 weeks ago
- ☆1,275Nov 20, 2025Updated 10 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios☆641Jun 12, 2026Updated 3 months ago
- OmniAgent: Audio-Guided Active Perception Agent for Omnimodal Audio-Video Understanding☆26Apr 9, 2026Updated 5 months ago
- 🔥 OneThinker: All-in-one Reasoning Model for Image and Video [CVPR 2026]☆468Feb 28, 2026Updated 6 months ago
- ☆23Jun 16, 2026Updated 3 months ago
- [ICLR 2026] Agentic Reinforced Policy Optimization (ARPO)☆1,124Sep 12, 2026Updated last week
- The repository of VG-Refiner paper☆20Dec 9, 2025Updated 9 months ago
- [ICLR 2026 Oral] Evaluating and Enhancing Multimodal Agentic Search with Structured Long Reasoning Chains☆39Apr 11, 2026Updated 5 months ago
- Reinforcement Learning via Self-Distillation (SDPO)☆1,103Jul 1, 2026Updated 2 months ago
- Explore the Multimodal “Aha Moment” on 2B Model☆623Mar 18, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Beyond SFT-to-RL: Pre-alignment via Black-BoxOn-Policy Distillation for Multimodal RL☆101May 6, 2026Updated 4 months ago
- Extend OpenRLHF to support LMM RL training for reproduction of DeepSeek-R1 on multimodal tasks.☆847May 14, 2025Updated last year
- Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL☆5,439Nov 13, 2025Updated 10 months ago
- The implementation for SIGIR 2026: Learning to Retrieve from Agent Trajectories.☆62Updated this week
- On Policy Distillation Build on top of Verl☆100Sep 3, 2026Updated 2 weeks ago
- ☆86May 8, 2026Updated 4 months ago
- Long-RL: Scaling RL to Long Sequences (NeurIPS 2025)☆729Sep 24, 2025Updated 11 months ago
- [ICLR 2026] The official repository for the paper "AdaReasoner: Dynamic Tool Orchestration for Iterative Visual Reasoning".☆84Aug 11, 2026Updated last month
- [ICML 2026] ZwZ model family: SOTA fine-grained perception performace; ZoomBench: a new challenging perception benchmark☆189May 4, 2026Updated 4 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [CVPR 2026] OpenMMReasoner: Pushing the Frontiers for Multimodal Reasoning with an Open and General Recipe☆166Mar 30, 2026Updated 5 months ago
- ✨✨ [ICLR 2026] Think Beyond Images☆585Sep 23, 2025Updated last year
- Unlocking Iterative Reasoning for Any Image Editor☆112Jan 18, 2026Updated 8 months ago
- Official codes of "Sketch-in-Latents: Eliciting Unified Reasoning in MLLMs"☆19Updated this week
- Official Repo of "Flow-OPD: On-Policy Distillation for Flow Matching Models"☆311Jun 24, 2026Updated 2 months ago
- Official Repository: A Comprehensive Benchmark for Logical Reasoning in MLLMs☆45Jun 17, 2025Updated last year
- GEMS: Agent-Native Multimodal Generation with Memory and Skills☆148Apr 1, 2026Updated 5 months ago