Official implementation of "Think, Then Verify: A Hypothesis–Verification Multi-Agent Framework for Long Video Understanding(CVPR'2026)"
☆28Aug 10, 2026Updated 2 weeks ago
Alternatives and similar repositories for VideoHV-Agent
Users that are interested in VideoHV-Agent are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆31Apr 9, 2026Updated 4 months ago
- The official repository of Omni-Weather. Code will be made publicly available soon.☆16Mar 30, 2026Updated 4 months ago
- [CVPR'2025] Narrating the Video: Boosting Text-Video Retrieval via Comprehensive Utilization of Frame-Level Captions☆19Jan 16, 2026Updated 7 months ago
- 基于C++ GUI Qt编写的HTTP在线音乐播放器☆10Dec 27, 2023Updated 2 years ago
- [CVPR 2026] LongVideo-R1: Smart Navigation for Low-cost Long Video Understanding☆52Jul 7, 2026Updated last month
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [NeurIPS 2025 Spotlight] Official PyTorch implementation of Vgent☆50Nov 30, 2025Updated 8 months ago
- [MICCAI2025] A latent motion profiling method for unsupervised cardiac phase detection.☆17Dec 15, 2025Updated 8 months ago
- This repository is the official Pytorch implementation of Balanced Product of Calibrated Experts for Long-Tailed Recognition (CVPR 2023).☆17Mar 13, 2025Updated last year
- [CVPR 2026] VideoSeek: Long-Horizon Video Agent with Tool-Guided Seeking☆68Mar 23, 2026Updated 5 months ago
- Official Code of "Random Parameter Pruning Attack (Accepeted by CVPR26)"☆16Feb 26, 2026Updated 5 months ago
- ☆20May 15, 2026Updated 3 months ago
- code of cvpr26 paper Symphony☆17Apr 7, 2026Updated 4 months ago
- ACM Multimedia 2023 (Oral) - RTQ: Rethinking Video-language Understanding Based on Image-text Model☆15Apr 7, 2026Updated 4 months ago
- Official Code for paper "Active Video Perception: Iterative Evidence Seeking for Agentic Long Video Understanding""☆19Jun 2, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ECCV 2026] StAR: Segment Anything Reasoner☆25Apr 2, 2026Updated 4 months ago
- Official implementation of "AgentRVOS: Reasoning Over Object Tracks for Zero-Shot Referring Video Object Segmentation".☆23Mar 25, 2026Updated 4 months ago
- ☆16Jul 17, 2026Updated last month
- [ACMMM'25] Referring Expression Instance Retrieval and A Strong End-to-End Baseline☆19Apr 7, 2026Updated 4 months ago
- The offical repo for "Play to the Score: Stage-Guided Dynamic Multi-Sensory Fusion for Robotic Manipulation", CoRL 2024 (ORAL)☆22Jun 25, 2025Updated last year
- [CVPR 2026] WISER: Wider Search, Deeper Thinking, and Adaptive Fusion for Training-Free Zero-Shot Composed Image Retrieval☆23Jun 17, 2026Updated 2 months ago
- VideoDetective: Clue Hunting via both Extrinsic Query and Intrinsic Relevance for Long Video Understanding☆58May 1, 2026Updated 3 months ago
- Code implementation of paper "MUSE: Mamba is Efficient Multi-scale Learner for Text-video Retrieval (AAAI2025)"☆26Feb 2, 2025Updated last year
- [ICLR 2026 Oral] Official Implementation of the paper "MetaEmbed: Scaling Multimodal Retrieval at Test-Time with Flexible Late Interactio…☆21Jul 2, 2026Updated last month
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆20Apr 21, 2026Updated 4 months ago
- [ECCV 2024] Official code repository of paper titled "Efficient 3D-Aware Facial Image Editing Via Attribute-Specific Prompt Learning"☆10Aug 2, 2024Updated 2 years ago
- Implementation of Poincare Embedding in PyTorch☆13Jul 27, 2017Updated 9 years ago
- ☆19Jun 22, 2025Updated last year
- [ICML 2026] HiDe: Rethinking The Zoom-IN method in High Resolution MLLMs via Hierarchical Decoupling☆30May 2, 2026Updated 3 months ago
- Cross-State Transition Attention Transformer for improved robotic manipulation with better temporal modeling; https://arxiv.org/abs/2510.…☆19Mar 8, 2026Updated 5 months ago
- Official Repository of ViTacGen: Robotic Pushing with Vision-to-Touch Generation (RA-L 2025 & ICRA 2026)☆18Feb 5, 2026Updated 6 months ago
- EVA: Efficient Reinforcement Learning for End-to-End Video Agent☆25May 6, 2026Updated 3 months ago
- ☆27Feb 25, 2026Updated 5 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ACMMM 2026] PLUME: Latent Reasoning Based Universal Multimodal Embedding☆25Apr 29, 2026Updated 3 months ago
- A Hierarchical Graph V-Net with Semi-supervised Pre-training for Breast Cancer Histology Image Classification" (IEEE TMI)☆22Oct 23, 2023Updated 2 years ago
- Official PyTorch implementation of "MM-PoisonRAG: Disrupting Multimodal RAG with Local and Global Poisoning Attacks"☆17Dec 4, 2025Updated 8 months ago
- ☆12Feb 7, 2018Updated 8 years ago
- CPLEX code of the E-VRPTW☆17Apr 8, 2024Updated 2 years ago
- PoseIt a multi-modal dataset that contains visual tactile data for holding poses☆15Feb 9, 2023Updated 3 years ago
- ☆26Jan 29, 2026Updated 6 months ago