Code implementation for the paper "Large-scale Pre-training for Grounded Video Caption Generation" (ICCV 2025)
☆33Jan 18, 2026Updated 7 months ago
Alternatives and similar repositories for grove
Users that are interested in grove are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation of "HowToCaption: Prompting LLMs to Transform Video Annotations at Scale." ECCV 2024☆59Aug 19, 2025Updated last year
- Strefer: Empowering Video LLMs with Space-Time Referring and Reasoning via Synthetic Instruction Data☆19Jun 2, 2026Updated 3 months ago
- ☆12Dec 6, 2024Updated last year
- B-cell Hybrid Immune Variant Engine☆12Jun 26, 2026Updated 2 months ago
- ☆20Jan 20, 2023Updated 3 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Implementation of "With a Little Help from my Temporal Context: Multimodal Egocentric Action Recognition, BMVC, 2021" in PyTorch☆20Dec 16, 2021Updated 4 years ago
- Library of models for Protein Function prediction (part of the 18th top solution out of 1625 teams in CAFA5)☆20May 23, 2025Updated last year
- Repository for the usage of Squidly.☆15Updated this week
- Official Code for the paper "UniversalVTG: A Univeral and Lightweight Foundation Model for Video Temporal Grounding"☆17Apr 15, 2026Updated 4 months ago
- ☆19Oct 28, 2025Updated 10 months ago
- Code for the paper "Learning to engineer protein flexibility".☆22Mar 24, 2026Updated 5 months ago
- ☆28Jul 18, 2025Updated last year
- [NeurIPS 2025] Panoptic Captioning: An Equivalence Bridge for Image and Text☆38Jan 31, 2026Updated 7 months ago
- Official repository for the paper "Evaluating variant effect prediction across viruses"☆22Mar 30, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICCV 2025] Object-centric Video Question Answering with Visual Grounding and Referring☆24Aug 8, 2025Updated last year
- C++/Python Library for Systematic Chemical Space Exploration☆25Dec 21, 2021Updated 4 years ago
- Code and data for the paper: Learning Action and Reasoning-Centric Image Editing from Videos and Simulation☆36Jun 30, 2025Updated last year
- High-Resolution Visual Reasoning via Multi-Turn Grounding-Based Reinforcement Learning☆56Jul 23, 2025Updated last year
- This is the offical repository of LLAVIDAL☆25Oct 4, 2025Updated 11 months ago
- CycleReward is a reward model trained on cycle consistency preferences to measure image-text alignment.☆57Nov 3, 2025Updated 10 months ago
- [EMNLP 2025 Oral] Official codebase for Seeing More, Saying More: Lightweight Language Experts are Dynamic Video Token Compressors.☆18Sep 7, 2025Updated last year
- Beyond Trivial Counterfactual Explanations with Diverse Valuable Explanations is a ServiceNow Research project that was started at Elemen…☆13Jul 31, 2023Updated 3 years ago
- [CVPR 2025] Official PyTorch code of "Enhancing Video-LLM Reasoning via Agent-of-Thoughts Distillation".☆59Jul 26, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Universal Video Temporal Grounding with Generative Multi-modal Large Language Models☆55May 20, 2026Updated 3 months ago
- Code implementation for our ECCV, 2022 paper titled "My View is the Best View: Procedure Learning from Egocentric Videos"☆35Feb 5, 2024Updated 2 years ago
- HT-Step is a large-scale article grounding dataset of temporal step annotations on how-to videos☆26Mar 20, 2024Updated 2 years ago
- Official PyTorch code of GroundVQA (CVPR'24)☆63Sep 13, 2024Updated last year
- The official code of "Thinking With Videos: Multimodal Tool-Augmented Reinforcement Learning for Long Video Reasoning"☆102Oct 15, 2025Updated 10 months ago
- A Spatial-Temporal Recurrent Neural Network for Video Saliency Prediction (TIP2021)☆13Jul 7, 2022Updated 4 years ago
- [CVPR 2024 Champions][ICLR 2025] Solutions for EgoVis Chanllenges in CVPR 2024☆136May 11, 2025Updated last year
- SpaceVLLM: Endowing Multimodal Large Language Model with Spatio-Temporal Video Grounding Capability☆17May 8, 2025Updated last year
- Codebase for the paper: "TIM: A Time Interval Machine for Audio-Visual Action Recognition"☆54Nov 7, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [CVPR 2023] OCTET: Object-aware Counterfactual Explanations☆19Dec 12, 2024Updated last year
- Official Code for the paper "HieraMamba: Video Temporal Grounding via Hierarchical Anchor-Mamba Pooling"☆17Apr 30, 2026Updated 4 months ago
- Code for the paper "GenHowTo: Learning to Generate Actions and State Transformations from Instructional Videos" published at CVPR 2024☆54Mar 3, 2024Updated 2 years ago
- Official Implementation (Pytorch) of the "VidChain: Chain-of-Tasks with Metric-based Direct Preference Optimization for Dense Video Capti…☆26Jan 26, 2025Updated last year
- [NIPS2025] VideoChat-R1 & R1.5: Enhancing Spatio-Temporal Perception and Reasoning via Reinforcement Fine-Tuning☆268Oct 18, 2025Updated 10 months ago
- ☆20Mar 3, 2025Updated last year
- ☆23Nov 4, 2024Updated last year