HyperEyes is a parallel multimodal search agent that fuses visual grounding and retrieval into a single atomic action, enabling concurrent search across multiple entities while treating inference efficiency as a first-class training objective.
β76May 23, 2026Updated 3 months ago
Alternatives and similar repositories for HyperEyes
Users that are interested in HyperEyes are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- πͺ Agent2World: Learning to Generate Symbolic World Models via Adaptive Multi-Agent Feedbackβ25Jan 29, 2026Updated 7 months ago
- Description for MV-MATHβ15Jul 20, 2025Updated last year
- The repository of VG-Refiner paperβ20Dec 9, 2025Updated 9 months ago
- Scaf-GRPO: Scaffolded Group Relative Policy Optimization for Enhancing LLM Reasoningβ26Feb 8, 2026Updated 7 months ago
- Rewards as Labels: Revisiting RLVR from a Classification Perspectiveβ25Jun 26, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- β85May 8, 2026Updated 4 months ago
- A Multimodal Reasoning Agent with Stateful Experiencesβ25Mar 31, 2026Updated 5 months ago
- π OpenSearch-VL provides a fully open recipe for training strong multimodal deep search agents through high-quality data curation, diverβ¦β282Jul 30, 2026Updated last month
- The official code of "Towards Long-horizon Agentic Multimodal Search"β30Apr 17, 2026Updated 4 months ago
- [NeurIPS'25] The official code of "PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning"β30Mar 30, 2026Updated 5 months ago
- [ICML 2026] ZwZ model family: SOTA fine-grained perception performace; ZoomBench: a new challenging perception benchmarkβ189May 4, 2026Updated 4 months ago
- [CVPR2026] Official codebase for the paper "Reasoning Within the Mind: Dynamic Multimodal Interleaving in Latent Space"β88May 12, 2026Updated 4 months ago
- Official data and code for the paper "VisBrowse-Bench: Benchmarking Visual-Native Search for Multimodal Browsing Agents".β14Mar 18, 2026Updated 5 months ago
- Official code, data, and models for "Hint Tuning: Less Data Makes Better Reasoners"β22Jul 8, 2026Updated 2 months ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- β48Jun 23, 2026Updated 2 months ago
- OmniGAIA: Towards Native Omni-Modal AI Agentsβ145Apr 2, 2026Updated 5 months ago
- β49Feb 9, 2026Updated 7 months ago
- β33Mar 17, 2026Updated 5 months ago
- [CVPR 2026] Official codes of "Monet: Reasoning in Latent Visual Space Beyond Image and Language"β223Mar 19, 2026Updated 5 months ago
- [ICLR 2026] The official repository for the paper "AdaReasoner: Dynamic Tool Orchestration for Iterative Visual Reasoning".β84Aug 11, 2026Updated last month
- The implementation for ACL 2026: MatchTIR: Fine-Grained Supervision for Tool-Integrated Reasoning via Bipartite Matching.β20Jul 27, 2026Updated last month
- Official codes of "Sketch-in-Latents: Eliciting Unified Reasoning in MLLMs"β19Feb 15, 2026Updated 6 months ago
- Benchmarking multimodal agents on realistic, ultra-challenging visual scenarios requiring long-horizon hybrid tool use.β74Mar 10, 2026Updated 6 months ago
- Proton VPN Special Offer - Get 70% off β’ AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- β37Apr 13, 2026Updated 5 months ago
- Code and data for the paper: AI Sees Your LocationβBut With A Bias Toward The Wealthy Worldβ19Dec 15, 2025Updated 8 months ago
- Paper writing guide for Zhuang Liu Lab @ Princeton Universityβ18Jun 24, 2026Updated 2 months ago
- An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scaleβ603Updated this week
- [CVPR 2026] Thinking with Programming Vision: Towards a Unified View for Thinking with Imagesβ74Jan 23, 2026Updated 7 months ago
- Search Self-Play: Pushing the Frontier of Agent Capability without Supervisionβ107Jul 23, 2026Updated last month
- [ICML 2026] XSkill: Continual Learning from Experience and Skills in Multimodal Agentsβ265May 13, 2026Updated 4 months ago
- [NeurIPS 2025] Reinforcing Spatial Reasoning in Vision-Language Models with Interwoven Thinking and Visual Drawingβ99Jul 27, 2025Updated last year
- ACL'2025: SoftCoT: Soft Chain-of-Thought for Efficient Reasoning with LLMs. and preprint: SoftCoT++: Test-Time Scaling with Soft Chain-ofβ¦β95May 30, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- https://huggingface.co/datasets/multimodal-reasoning-lab/Zebra-CoTβ140Jan 30, 2026Updated 7 months ago
- β27Mar 17, 2026Updated 5 months ago
- Agentic MLLMsβ215Oct 24, 2025Updated 10 months ago
- β88Jun 16, 2026Updated 2 months ago
- [ECCV 26'] Official codebase for the paper LaViTβ35Jul 30, 2026Updated last month
- Vision-OPD is a regional-to-global on-policy self-distillation framework that transfers a model's own privileged crop-conditioned perceptβ¦β308Jul 17, 2026Updated last month
- [CVPR 2026] OpenMMReasoner: Pushing the Frontiers for Multimodal Reasoning with an Open and General Recipeβ166Mar 30, 2026Updated 5 months ago