☆16Jun 2, 2025Updated last year
Alternatives and similar repositories for CTVR
Users that are interested in CTVR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆24Jul 23, 2025Updated last year
- [CVPR2025] Official code for Lost in Translation Found in Context☆24Jan 14, 2026Updated 7 months ago
- Video-Language Continual Learning Benchmark☆20Oct 30, 2024Updated last year
- UCPM: Uncertainty-Guided Cross-Modal Retrieval with Partially Mismatched Pairs (TIP 2025, pytorch code)☆25Apr 16, 2026Updated 4 months ago
- [CVPR'25] ConText-CIR: Learning from Concepts in Text for Composed Image Retrieval☆16Jun 17, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆13Apr 12, 2026Updated 4 months ago
- [ICCV2021] Official Pytorch implementation for SDGZSL (Semantics Disentangling for Generalized Zero-Shot Learning)☆41Dec 16, 2021Updated 4 years ago
- [CVPR 2026 main] Imagine Before Concentration: Diffusion-Guided Registers Enhance Partially Relevant Video Retrieval☆37Jun 1, 2026Updated 2 months ago
- Agentic Keyframe Search for Video Question Answering☆18Jun 30, 2026Updated 2 months ago
- Official Implementation of ISR-DPO:Aligning Large Multimodal Models for Videos by Iterative Self-Retrospective DPO (AAAI'25)☆23Nov 25, 2025Updated 9 months ago
- The Source Code for IF-VidCap @ICLR 2026☆18Oct 22, 2025Updated 10 months ago
- [ICCV 25] Official repository of "Collaborative Instance Object Navigation: Leveraging Uncertainty-Awareness to Minimize Human-Agent Dial…☆32Apr 1, 2026Updated 4 months ago
- Text Proxy: Decomposing Retrieval from a 1-to-N Relationship into N 1-to-1 Relationships for Text-Video Retrieval -- AAAI2025☆22May 8, 2026Updated 3 months ago
- This repository is an official implementation of AVIGATE (CVPR 2025, oral)☆19Aug 21, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆16Aug 28, 2024Updated 2 years ago
- A paper list of partially relevant video retrieval☆45Jul 21, 2026Updated last month
- Learning Situation Hyper-Graphs for Video Question Answering☆23Feb 16, 2024Updated 2 years ago
- ☆19Jul 28, 2025Updated last year
- 大连理工大学编译原理课程设计☆10Jan 1, 2024Updated 2 years ago
- We introduce the direct document relevance optimization (DDRO) for training a pairwise ranker model. DDRO encourages the model to focus o…☆40Jul 2, 2026Updated last month
- ☆17Apr 21, 2026Updated 4 months ago
- OpenDLM is an open-source library focused on sampling algorithms for Diffusion Language Models (DLMs).☆15Aug 5, 2025Updated last year
- This is an official PyTorch Implementation of Neighbor Relations Matter in Video Scene Detection.☆30Mar 19, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Awesome Compositional Zero-shot Learning papers.☆15Aug 26, 2025Updated last year
- [AAAI'25]: Building a Multi-modal Spatiotemporal Expert for Zero-shot Action Recognition with CLIP☆23Aug 5, 2025Updated last year
- The official implementation of MaskGRPO: Consolidating Reinforcement Learning for Multimodal Discrete Diffusion Models. (ICLR 2026, arxiv…☆19Jan 27, 2026Updated 7 months ago
- [CVPR 2025] CL-MoE: Enhancing Multimodal Large Language Model with Dual Momentum Mixture-of-Experts for Continual Visual Question Answeri…☆62Jun 16, 2025Updated last year
- A curated list of awesome succulent and cactus identification, cultivation, care, and advisory resources.☆11Aug 28, 2016Updated 10 years ago
- Official Pytorch Implementation of 'BAM-DETR: Boundary-Aligned Moment Detection Transformer for Temporal Sentence Grounding in Videos'☆36Feb 26, 2025Updated last year
- Grounded Visual Token Sampling (GroundVTS), a Vid-LLM architecture designed to enhance VTG performance through adaptive and efficient vis…☆16Jun 12, 2026Updated 2 months ago
- [NeurIPS25] Official Implementation (Pytorch) of "DeepVideo-R1"☆38Feb 22, 2026Updated 6 months ago
- CoFiRec: Coarse-to-Fine Tokenization for Generative Recommendationn☆21Jan 23, 2026Updated 7 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [ICCV'25] HERMES: temporal-coHERent long-forM understanding with Episodes and Semantics☆37Sep 10, 2025Updated 11 months ago
- Preprocessed data of SignDiff: Learning Diffusion Models for American Sign Language Production☆19May 1, 2025Updated last year
- Code for our works: LCSA, C2SLR, and SRM☆22Nov 22, 2024Updated last year
- [ICLR 2025] This repo is the official implementation of our paper "Learning Fine-Grained Representations through Textual Token Disentangl…☆23Jul 28, 2025Updated last year
- ☆13Jan 3, 2025Updated last year
- official PyTorch implementation of paper "Adversarial Bipartite Graph Learning for Video Domain Adaptation" (MM2020 Oral)☆11Jun 16, 2022Updated 4 years ago
- Source code for paper "Targeted Adversarial Attack for Deep Cross-modal Hashing Retrieval".☆20May 4, 2023Updated 3 years ago