☆16Jun 2, 2025Updated last year
Alternatives and similar repositories for CTVR
Users that are interested in CTVR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Note: DO NOT USE IT! THIS CODE IS PROVEN TO CONTAIN DATA LEAKAGE! Archive version of "Text Is MASS: Modeling as Stochastic Embedding for …☆23May 1, 2025Updated last year
- ☆24Jul 23, 2025Updated last year
- [CVPR2025] Official code for Lost in Translation Found in Context☆24Jan 14, 2026Updated 8 months ago
- Video-Language Continual Learning Benchmark☆20Oct 30, 2024Updated last year
- UCPM: Uncertainty-Guided Cross-Modal Retrieval with Partially Mismatched Pairs (TIP 2025, pytorch code)☆25Apr 16, 2026Updated 5 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆13Apr 12, 2026Updated 5 months ago
- [ICCV2021] Official Pytorch implementation for SDGZSL (Semantics Disentangling for Generalized Zero-Shot Learning)☆42Dec 16, 2021Updated 4 years ago
- [CVPR 2026 main] Imagine Before Concentration: Diffusion-Guided Registers Enhance Partially Relevant Video Retrieval☆39Jun 1, 2026Updated 3 months ago
- Agentic Keyframe Search for Video Question Answering☆18Jun 30, 2026Updated 2 months ago
- Official Implementation of ISR-DPO:Aligning Large Multimodal Models for Videos by Iterative Self-Retrospective DPO (AAAI'25)☆23Nov 25, 2025Updated 9 months ago
- The Source Code for IF-VidCap @ICLR 2026☆18Oct 22, 2025Updated 10 months ago
- [ICCV 25] Official repository of "Collaborative Instance Object Navigation: Leveraging Uncertainty-Awareness to Minimize Human-Agent Dial…☆32Apr 1, 2026Updated 5 months ago
- Text Proxy: Decomposing Retrieval from a 1-to-N Relationship into N 1-to-1 Relationships for Text-Video Retrieval -- AAAI2025☆22May 8, 2026Updated 4 months ago
- This repository is an official implementation of AVIGATE (CVPR 2025, oral)☆19Aug 21, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆16Aug 28, 2024Updated 2 years ago
- A paper list of partially relevant video retrieval☆45Sep 12, 2026Updated last week
- Learning Situation Hyper-Graphs for Video Question Answering☆23Feb 16, 2024Updated 2 years ago
- ☆19Jul 28, 2025Updated last year
- 大连理工大学编译原理课程设计☆10Jan 1, 2024Updated 2 years ago
- The Source Code for MT-Video-Bench @ ACL Findings 2026☆21Jan 20, 2026Updated 7 months ago
- We introduce the direct document relevance optimization (DDRO) for training a pairwise ranker model. DDRO encourages the model to focus o…☆40Jul 2, 2026Updated 2 months ago
- ☆16Apr 21, 2026Updated 4 months ago
- OpenDLM is an open-source library focused on sampling algorithms for Diffusion Language Models (DLMs).☆15Aug 5, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- This is an official PyTorch Implementation of Neighbor Relations Matter in Video Scene Detection.☆30Mar 19, 2025Updated last year
- [AAAI'25]: Building a Multi-modal Spatiotemporal Expert for Zero-shot Action Recognition with CLIP☆23Aug 5, 2025Updated last year
- ☆15Jun 5, 2025Updated last year
- The official implementation of MaskGRPO: Consolidating Reinforcement Learning for Multimodal Discrete Diffusion Models. (ICLR 2026, arxiv…☆19Jan 27, 2026Updated 7 months ago
- [CVPR 2025] CL-MoE: Enhancing Multimodal Large Language Model with Dual Momentum Mixture-of-Experts for Continual Visual Question Answeri…☆62Jun 16, 2025Updated last year
- Official Pytorch Implementation of 'BAM-DETR: Boundary-Aligned Moment Detection Transformer for Temporal Sentence Grounding in Videos'☆36Feb 26, 2025Updated last year
- [NeurIPS25] Official Implementation (Pytorch) of "DeepVideo-R1"☆38Feb 22, 2026Updated 6 months ago
- CoFiRec: Coarse-to-Fine Tokenization for Generative Recommendationn☆22Jan 23, 2026Updated 7 months ago
- [ICCV'25] HERMES: temporal-coHERent long-forM understanding with Episodes and Semantics☆37Sep 10, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Preprocessed data of SignDiff: Learning Diffusion Models for American Sign Language Production☆19May 1, 2025Updated last year
- Project for SNARE benchmark☆11Jun 5, 2024Updated 2 years ago
- Code for our works: LCSA, C2SLR, and SRM☆22Nov 22, 2024Updated last year
- [ICLR 2025] This repo is the official implementation of our paper "Learning Fine-Grained Representations through Textual Token Disentangl…☆23Jul 28, 2025Updated last year
- Code for "Multi-level Relevance Document Identifier Learning for Generative Retrieval". ACL 2025.☆24Nov 5, 2025Updated 10 months ago
- ☆13Jan 3, 2025Updated last year
- Source code for paper "Targeted Adversarial Attack for Deep Cross-modal Hashing Retrieval".☆20May 4, 2023Updated 3 years ago