HyperEyes is a parallel multimodal search agent that fuses visual grounding and retrieval into a single atomic action, enabling concurrent search across multiple entities while treating inference efficiency as a first-class training objective.
β74May 23, 2026Updated 3 months ago
Alternatives and similar repositories for HyperEyes
Users that are interested in HyperEyes are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- πͺ Agent2World: Learning to Generate Symbolic World Models via Adaptive Multi-Agent Feedbackβ23Jan 29, 2026Updated 6 months ago
- Description for MV-MATHβ15Jul 20, 2025Updated last year
- The repository of VG-Refiner paperβ20Dec 9, 2025Updated 8 months ago
- Scaf-GRPO: Scaffolded Group Relative Policy Optimization for Enhancing LLM Reasoningβ25Feb 8, 2026Updated 6 months ago
- Rewards as Labels: Revisiting RLVR from a Classification Perspectiveβ25Jun 26, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A Multimodal Reasoning Agent with Stateful Experiencesβ25Mar 31, 2026Updated 4 months ago
- β80May 8, 2026Updated 3 months ago
- π OpenSearch-VL provides a fully open recipe for training strong multimodal deep search agents through high-quality data curation, diverβ¦β270Jul 30, 2026Updated 3 weeks ago
- The official code of "Towards Long-horizon Agentic Multimodal Search"β29Apr 17, 2026Updated 4 months ago
- [NeurIPS'25] The official code of "PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning"β30Mar 30, 2026Updated 4 months ago
- [ICML 2026] ZwZ model family: SOTA fine-grained perception performace; ZoomBench: a new challenging perception benchmarkβ187May 4, 2026Updated 3 months ago
- [CVPR2026] Official codebase for the paper "Reasoning Within the Mind: Dynamic Multimodal Interleaving in Latent Space"β88May 12, 2026Updated 3 months ago
- Official data and code for the paper "VisBrowse-Bench: Benchmarking Visual-Native Search for Multimodal Browsing Agents".β15Mar 18, 2026Updated 5 months ago
- β47Jun 23, 2026Updated 2 months ago
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Official code, data, and models for "Hint Tuning: Less Data Makes Better Reasoners"β22Jul 8, 2026Updated last month
- OmniGAIA: Towards Native Omni-Modal AI Agentsβ145Apr 2, 2026Updated 4 months ago
- β48Feb 9, 2026Updated 6 months ago
- β32Mar 17, 2026Updated 5 months ago
- Paper writing guide for Zhuang Liu Lab @ Princeton Universityβ17Jun 24, 2026Updated 2 months ago
- [CVPR 2026] Official codes of "Monet: Reasoning in Latent Visual Space Beyond Image and Language"β219Mar 19, 2026Updated 5 months ago
- [ICLR 2026] The official repository for the paper "AdaReasoner: Dynamic Tool Orchestration for Iterative Visual Reasoning".β82Aug 11, 2026Updated last week
- The implementation for ACL 2026: MatchTIR: Fine-Grained Supervision for Tool-Integrated Reasoning via Bipartite Matching.β20Jul 27, 2026Updated 3 weeks ago
- Official codes of "Sketch-in-Latents: Eliciting Unified Reasoning in MLLMs"β17Feb 15, 2026Updated 6 months ago
- Proton VPN Special Offer - Get 70% off β’ AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Benchmarking multimodal agents on realistic, ultra-challenging visual scenarios requiring long-horizon hybrid tool use.β72Mar 10, 2026Updated 5 months ago
- β37Apr 13, 2026Updated 4 months ago
- Code and data for the paper: AI Sees Your LocationβBut With A Bias Toward The Wealthy Worldβ19Dec 15, 2025Updated 8 months ago
- An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scaleβ578Updated this week
- [CVPR 2026] Thinking with Programming Vision: Towards a Unified View for Thinking with Imagesβ72Jan 23, 2026Updated 7 months ago
- Search Self-Play: Pushing the Frontier of Agent Capability without Supervisionβ106Jul 23, 2026Updated last month
- [ICML 2026] XSkill: Continual Learning from Experience and Skills in Multimodal Agentsβ255May 13, 2026Updated 3 months ago
- ε¦ζ―δΈ»ι‘΅ | Academic Pageβ14Updated this week
- ACL'2025: SoftCoT: Soft Chain-of-Thought for Efficient Reasoning with LLMs. and preprint: SoftCoT++: Test-Time Scaling with Soft Chain-ofβ¦β95May 30, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- https://huggingface.co/datasets/multimodal-reasoning-lab/Zebra-CoTβ139Jan 30, 2026Updated 6 months ago
- β27Mar 17, 2026Updated 5 months ago
- Agentic MLLMsβ216Oct 24, 2025Updated 10 months ago
- β88Jun 16, 2026Updated 2 months ago
- [ECCV 26'] Official codebase for the paper LaViTβ35Jul 30, 2026Updated 3 weeks ago
- Vision-OPD is a regional-to-global on-policy self-distillation framework that transfers a model's own privileged crop-conditioned perceptβ¦β287Jul 17, 2026Updated last month
- [CVPR 2026] OpenMMReasoner: Pushing the Frontiers for Multimodal Reasoning with an Open and General Recipeβ166Mar 30, 2026Updated 4 months ago