HyperEyes is a parallel multimodal search agent that fuses visual grounding and retrieval into a single atomic action, enabling concurrent search across multiple entities while treating inference efficiency as a first-class training objective.
β70May 23, 2026Updated 2 months ago
Alternatives and similar repositories for HyperEyes
Users that are interested in HyperEyes are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- πͺ Agent2World: Learning to Generate Symbolic World Models via Adaptive Multi-Agent Feedbackβ23Jan 29, 2026Updated 6 months ago
- The repository of VG-Refiner paperβ20Dec 9, 2025Updated 7 months ago
- Scaf-GRPO: Scaffolded Group Relative Policy Optimization for Enhancing LLM Reasoningβ24Feb 8, 2026Updated 5 months ago
- A Multimodal Reasoning Agent with Stateful Experiencesβ24Mar 31, 2026Updated 4 months ago
- β77May 8, 2026Updated 2 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- π OpenSearch-VL provides a fully open recipe for training strong multimodal deep search agents through high-quality data curation, diverβ¦β260Updated this week
- The official code of "Towards Long-horizon Agentic Multimodal Search"β28Apr 17, 2026Updated 3 months ago
- [NeurIPS'25] The official code of "PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning"β30Mar 30, 2026Updated 4 months ago
- [ICML 2026] ZwZ model family: SOTA fine-grained perception performace; ZoomBench: a new challenging perception benchmarkβ183May 4, 2026Updated 2 months ago
- [CVPR2026] Official codebase for the paper "Reasoning Within the Mind: Dynamic Multimodal Interleaving in Latent Space"β85May 12, 2026Updated 2 months ago
- Official data and code for the paper "VisBrowse-Bench: Benchmarking Visual-Native Search for Multimodal Browsing Agents".β15Mar 18, 2026Updated 4 months ago
- β46Jun 23, 2026Updated last month
- OmniGAIA: Towards Native Omni-Modal AI Agentsβ142Apr 2, 2026Updated 4 months ago
- β47Feb 9, 2026Updated 5 months ago
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- β32Mar 17, 2026Updated 4 months ago
- Paper writing guide for Zhuang Liu Lab @ Princeton Universityβ16Jun 24, 2026Updated last month
- [CVPR 2026] Official codes of "Monet: Reasoning in Latent Visual Space Beyond Image and Language"β216Mar 19, 2026Updated 4 months ago
- [ICLR 2026] The official repository for the paper "AdaReasoner: Dynamic Tool Orchestration for Iterative Visual Reasoning".β82Feb 27, 2026Updated 5 months ago
- The implementation for ACL 2026: MatchTIR: Fine-Grained Supervision for Tool-Integrated Reasoning via Bipartite Matching.β20Jul 27, 2026Updated last week
- Official codes of "Sketch-in-Latents: Eliciting Unified Reasoning in MLLMs"β17Feb 15, 2026Updated 5 months ago
- Benchmarking multimodal agents on realistic, ultra-challenging visual scenarios requiring long-horizon hybrid tool use.β70Mar 10, 2026Updated 4 months ago
- β36Apr 13, 2026Updated 3 months ago
- [CVPR 2026] Thinking with Programming Vision: Towards a Unified View for Thinking with Imagesβ72Jan 23, 2026Updated 6 months ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Search Self-Play: Pushing the Frontier of Agent Capability without Supervisionβ105Jul 23, 2026Updated last week
- [NeurIPS 2025] Reinforcing Spatial Reasoning in Vision-Language Models with Interwoven Thinking and Visual Drawingβ97Jul 27, 2025Updated last year
- [ICML 2026] XSkill: Continual Learning from Experience and Skills in Multimodal Agentsβ243May 13, 2026Updated 2 months ago
- ε¦ζ―δΈ»ι‘΅ | Academic Pageβ14Updated this week
- ACL'2025: SoftCoT: Soft Chain-of-Thought for Efficient Reasoning with LLMs. and preprint: SoftCoT++: Test-Time Scaling with Soft Chain-ofβ¦β94May 30, 2025Updated last year
- https://huggingface.co/datasets/multimodal-reasoning-lab/Zebra-CoTβ137Jan 30, 2026Updated 6 months ago
- β27Mar 17, 2026Updated 4 months ago
- Agentic MLLMsβ216Oct 24, 2025Updated 9 months ago
- β84Jun 16, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ECCV 26'] Official codebase for the paper LaViTβ35Updated this week
- Vision-OPD is a regional-to-global on-policy self-distillation framework that transfers a model's own privileged crop-conditioned perceptβ¦β239Jul 17, 2026Updated 2 weeks ago
- [CVPR 2026] OpenMMReasoner: Pushing the Frontiers for Multimodal Reasoning with an Open and General Recipeβ165Mar 30, 2026Updated 4 months ago
- This repository contains the code for the paper βNeuro-Symbolic Query Compilerβ, accepted to the Findings of ACL 2025.β18Oct 20, 2025Updated 9 months ago
- Official code for "Self-Distilled Agentic Reinforcement Learning"β318Updated this week
- [ICML 2025] Official code of "DAMA: Data- and Model-aware Alignment of Multi-modal LLMs"β16May 24, 2025Updated last year
- β35Apr 10, 2026Updated 3 months ago