Official implementation of "Think, Then Verify: A Hypothesis–Verification Multi-Agent Framework for Long Video Understanding(CVPR'2026)"
☆29Aug 10, 2026Updated last month
Alternatives and similar repositories for VideoHV-Agent
Users that are interested in VideoHV-Agent are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆34Apr 9, 2026Updated 5 months ago
- The official repository of Omni-Weather. Code will be made publicly available soon.☆16Mar 30, 2026Updated 5 months ago
- 基于C++ GUI Qt编写的HTTP在线音乐播放器☆10Dec 27, 2023Updated 2 years ago
- [CVPR 2026] LongVideo-R1: Smart Navigation for Low-cost Long Video Understanding☆54Jul 7, 2026Updated 2 months ago
- 街道社区活动小程序是一款基于微信小程序平台的应用程序,旨在为社区居民提供便捷的社区活动信息查询、报名、参与等服务。通过该小程序,居民可以更加便捷地参与社区活动,增强社区凝聚力和归属感,促进社区和谐发展。 该小程序主要包括以下功能: 社区活动信息浏览:居民可以通过小程序浏览…☆14Oct 4, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [NeurIPS 2025 Spotlight] Official PyTorch implementation of Vgent☆50Nov 30, 2025Updated 9 months ago
- 基于TensorFlow官方Android端示例,对选择的图片进行风格迁移☆11Apr 9, 2017Updated 9 years ago
- [MICCAI2025] A latent motion profiling method for unsupervised cardiac phase detection.☆17Dec 15, 2025Updated 8 months ago
- Official Code of "Random Parameter Pruning Attack (Accepeted by CVPR26)"☆17Feb 26, 2026Updated 6 months ago
- Visual Speech Recongnition☆22Dec 24, 2024Updated last year
- ☆20May 15, 2026Updated 3 months ago
- ACM Multimedia 2023 (Oral) - RTQ: Rethinking Video-language Understanding Based on Image-text Model☆15Apr 7, 2026Updated 5 months ago
- Official Code for paper "Active Video Perception: Iterative Evidence Seeking for Agentic Long Video Understanding""☆20Jun 2, 2026Updated 3 months ago
- [ECCV 2026] StAR: Segment Anything Reasoner☆25Apr 2, 2026Updated 5 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official implementation of "AgentRVOS: Reasoning Over Object Tracks for Zero-Shot Referring Video Object Segmentation".☆23Mar 25, 2026Updated 5 months ago
- ☆16Jul 17, 2026Updated last month
- [ACMMM'25] Referring Expression Instance Retrieval and A Strong End-to-End Baseline☆19Apr 7, 2026Updated 5 months ago
- ☆21Jan 17, 2025Updated last year
- [ICLR-2026] Official Implementation of our paper "THOR: Tool-Integrated Hierarchical Optimization via RL for Mathematical Reasoning".☆32Aug 4, 2026Updated last month
- ☆17Aug 11, 2023Updated 3 years ago
- [CVPR 2026] WISER: Wider Search, Deeper Thinking, and Adaptive Fusion for Training-Free Zero-Shot Composed Image Retrieval☆25Jun 17, 2026Updated 2 months ago
- VideoDetective: Clue Hunting via both Extrinsic Query and Intrinsic Relevance for Long Video Understanding☆58May 1, 2026Updated 4 months ago
- ☆26Apr 7, 2025Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- [ICLR 2026 Oral] Official Implementation of the paper "MetaEmbed: Scaling Multimodal Retrieval at Test-Time with Flexible Late Interactio…☆22Jul 2, 2026Updated 2 months ago
- MoDem-V2 combines the sample efficiency of the original MoDem with conservative exploration in order to quickly and safely learn manipula…☆25Apr 1, 2024Updated 2 years ago
- [ECCV 2024] Official code repository of paper titled "Efficient 3D-Aware Facial Image Editing Via Attribute-Specific Prompt Learning"☆10Aug 2, 2024Updated 2 years ago
- ☆19Jun 22, 2025Updated last year
- ☆21Apr 21, 2026Updated 4 months ago
- Cross-State Transition Attention Transformer for improved robotic manipulation with better temporal modeling; https://arxiv.org/abs/2510.…☆19Mar 8, 2026Updated 6 months ago
- ☆12Jul 14, 2022Updated 4 years ago
- A curated list of resources in audio visual question answering and related area. :-)☆18Jun 29, 2025Updated last year
- The code repository for "Cross-Modal and Hierarchical Modeling of Video and Text" in PyTorch☆20Apr 26, 2020Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆20Feb 23, 2026Updated 6 months ago
- [ACMMM 2026] PLUME: Latent Reasoning Based Universal Multimodal Embedding☆25Apr 29, 2026Updated 4 months ago
- ☆29Jan 13, 2024Updated 2 years ago
- A Hierarchical Graph V-Net with Semi-supervised Pre-training for Breast Cancer Histology Image Classification" (IEEE TMI)☆22Oct 23, 2023Updated 2 years ago
- ☆30May 18, 2025Updated last year
- Code release for the paper "Progress-Aware Video Frame Captioning" (CVPR 2025)☆26Jul 16, 2025Updated last year
- ☆12Feb 7, 2018Updated 8 years ago