☆71Jan 4, 2026Updated 7 months ago
Alternatives and similar repositories for MM-BrowseComp
Users that are interested in MM-BrowseComp are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MMDeepResearch-Bench (MMDR)☆32Aug 10, 2026Updated 3 weeks ago
- COLM2026☆38Jul 9, 2026Updated last month
- ☆16Nov 1, 2025Updated 9 months ago
- MMhops-R1: Multimodal Multi-hop Reasoning☆16Aug 17, 2026Updated 2 weeks ago
- Official code repo of Video-Browser: Towards Agentic Open-web Video Browsing☆28Jan 19, 2026Updated 7 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The Source Code for WebCompass☆25May 2, 2026Updated 3 months ago
- [ACL-2026] MMSearch-R1 is an end-to-end RL framework that enables LMMs to perform on-demand, multi-turn search with real-world multimodal…☆483Apr 7, 2026Updated 4 months ago
- [ICLR'2026] R-HORIZON: How Far Can Your Large Reasoning Model Really Go in Breadth and Depth?☆18Oct 21, 2025Updated 10 months ago
- RAG-RewardBench: Benchmarking Reward Models in Retrieval Augmented Generation for Preference Alignment☆18Dec 19, 2024Updated last year
- Data Synthesis for Deep Research Based on Semi-Structured Data☆216Jul 14, 2026Updated last month
- A Comprehensive Survey on Evaluating Reasoning Capabilities in Multimodal Large Language Models.☆76Mar 18, 2025Updated last year
- The official repository of "R-4B: Incentivizing General-Purpose Auto-Thinking Capability in MLLMs via Bi-Mode Integration"☆141Sep 4, 2025Updated 11 months ago
- ☆220Dec 19, 2025Updated 8 months ago
- Implementation for "The Scalability of Simplicity: Empirical Analysis of Vision-Language Learning with a Single Transformer"☆85Oct 29, 2025Updated 10 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Scaling Deep Research via Reinforcement Learning in Real-world Environments.☆799May 10, 2026Updated 3 months ago
- "DeepResearch-Eval: An End-to-End Evaluation Framework for DeepResearch Systems"☆50Oct 16, 2025Updated 10 months ago
- Official Code for "Mini-o3: Scaling Up Reasoning Patterns and Interaction Turns for Visual Search"☆425Jan 29, 2026Updated 7 months ago
- ☆46Dec 30, 2024Updated last year
- ☆12Mar 12, 2023Updated 3 years ago
- Code for DVD A Diagnostic Dataset for Multi-step Reasoning in Video Grounded Dialogue☆14Oct 12, 2021Updated 4 years ago
- An Open-Source Large-Scale Reinforcement Learning Project for Search Agents☆609Nov 26, 2025Updated 9 months ago
- This repo contains evaluation code for the paper "AV-Odyssey: Can Your Multimodal LLMs Really Understand Audio-Visual Information?"☆31Dec 23, 2024Updated last year
- Extending context length of visual language models☆12Dec 18, 2024Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [arXiv 2025] SAGE: Training Smart Any-Horizon Agents for Long Video Reasoning with Reinforcement Learning☆70Dec 17, 2025Updated 8 months ago
- ☆48Aug 15, 2023Updated 3 years ago
- ☆18Jun 18, 2024Updated 2 years ago
- OpenThinkIMG is an end-to-end open-source framework that empowers LVLMs to think with images.☆403Jun 1, 2025Updated last year
- The official repo of "WebExplorer: Explore and Evolve for Training Long-Horizon Web Agents"☆121Sep 29, 2025Updated 11 months ago
- [TPAMI 2026] Ego-R1: Agentic Chain-of-Tool-Thought for Ultra-Long Egocentric Video Reasoning☆168Jun 10, 2026Updated 2 months ago
- [ICLR 2026] End-to-End Reinforcement Learning for Multi-Turn Tool-Integrated Reasoning☆405Mar 30, 2026Updated 5 months ago
- Beyond Entities: A Large-Scale Multi-Modal Knowledge Graph with Triplet Fact Grounding☆11May 23, 2024Updated 2 years ago
- ☆192Aug 13, 2026Updated 2 weeks ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Long-RL: Scaling RL to Long Sequences (NeurIPS 2025)☆729Sep 24, 2025Updated 11 months ago
- Math-VR Benchmark & CodePlot-CoT: Mathematical Visual Reasoning by Thinking with Code-Driven Images☆64Nov 4, 2025Updated 9 months ago
- ☆18Dec 25, 2021Updated 4 years ago
- [ECCV 2024] Learning Video Context as Interleaved Multimodal Sequences☆46Mar 11, 2025Updated last year
- Official code of *Virgo: A Preliminary Exploration on Reproducing o1-like MLLM*☆20May 27, 2025Updated last year
- ☆16Mar 1, 2026Updated 5 months ago
- MVU-Eval @NeurIPS DB 2025☆16Nov 11, 2025Updated 9 months ago