☆16Jun 2, 2025Updated last year
Alternatives and similar repositories for CTVR
Users that are interested in CTVR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Note: DO NOT USE IT! THIS CODE IS PROVEN TO CONTAIN DATA LEAKAGE! Archive version of "Text Is MASS: Modeling as Stochastic Embedding for …☆23May 1, 2025Updated last year
- ☆23Jul 23, 2025Updated 11 months ago
- [CVPR2025] Official code for Lost in Translation Found in Context☆24Jan 14, 2026Updated 6 months ago
- Video-Language Continual Learning Benchmark☆20Oct 30, 2024Updated last year
- UCPM: Uncertainty-Guided Cross-Modal Retrieval with Partially Mismatched Pairs (TIP 2025, pytorch code)☆25Apr 16, 2026Updated 3 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official PyTorch implementation of our paper "Dispersing Prompt Expansion for Class-Agnostic Object Detection" (NeurIPS 2024)☆14Jan 19, 2025Updated last year
- [CVPR'25] ConText-CIR: Learning from Concepts in Text for Composed Image Retrieval☆16Jun 17, 2026Updated last month
- [CVPR 2026 main] Imagine Before Concentration: Diffusion-Guided Registers Enhance Partially Relevant Video Retrieval☆37Jun 1, 2026Updated last month
- [ICCV2021] Official Pytorch implementation for SDGZSL (Semantics Disentangling for Generalized Zero-Shot Learning)☆40Dec 16, 2021Updated 4 years ago
- Official Implementation of ISR-DPO:Aligning Large Multimodal Models for Videos by Iterative Self-Retrospective DPO (AAAI'25)☆23Nov 25, 2025Updated 7 months ago
- The Source Code for IF-VidCap @ICLR 2026☆19Oct 22, 2025Updated 8 months ago
- Text Proxy: Decomposing Retrieval from a 1-to-N Relationship into N 1-to-1 Relationships for Text-Video Retrieval -- AAAI2025☆21May 8, 2026Updated 2 months ago
- This repository is an official implementation of AVIGATE (CVPR 2025, oral)☆19Aug 21, 2025Updated 10 months ago
- 将训练好的人脸分类器模型文件转换为.pb格式,促进工程应用。☆11Jan 1, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A paper list of partially relevant video retrieval☆42Jun 1, 2026Updated last month
- ☆16Aug 28, 2024Updated last year
- Learning Situation Hyper-Graphs for Video Question Answering☆23Feb 16, 2024Updated 2 years ago
- ☆19Jul 28, 2025Updated 11 months ago
- 大连理工大学编译原理课程设计☆10Jan 1, 2024Updated 2 years ago
- The Source Code for MT-Video-Bench @ ACL Findings 2026☆22Jan 20, 2026Updated 6 months ago
- We introduce the direct document relevance optimization (DDRO) for training a pairwise ranker model. DDRO encourages the model to focus o…☆39Jul 2, 2026Updated 2 weeks ago
- [NeurIPS'24] I2EBench: A Comprehensive Benchmark for Instruction-based Image Editing☆35Dec 9, 2025Updated 7 months ago
- ☆15Apr 21, 2026Updated 3 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- This is an official PyTorch Implementation of Neighbor Relations Matter in Video Scene Detection.☆29Mar 19, 2025Updated last year
- OpenDLM is an open-source library focused on sampling algorithms for Diffusion Language Models (DLMs).☆15Aug 5, 2025Updated 11 months ago
- [AAAI'25]: Building a Multi-modal Spatiotemporal Expert for Zero-shot Action Recognition with CLIP☆23Aug 5, 2025Updated 11 months ago
- ☆15Jun 5, 2025Updated last year
- [CVPR 2025] CL-MoE: Enhancing Multimodal Large Language Model with Dual Momentum Mixture-of-Experts for Continual Visual Question Answeri…☆60Jun 16, 2025Updated last year
- The official implementation of MaskGRPO: Consolidating Reinforcement Learning for Multimodal Discrete Diffusion Models. (ICLR 2026, arxiv…☆19Jan 27, 2026Updated 5 months ago
- 🏆 🥇 Winner Solution for ICCV VQualA 2025 EVQA-SnapUGC Challenge at VQualA 2025 Workshop @ ICCV 2025☆18Aug 8, 2025Updated 11 months ago
- Grounded Visual Token Sampling (GroundVTS), a Vid-LLM architecture designed to enhance VTG performance through adaptive and efficient vis…☆16Jun 12, 2026Updated last month
- Official Pytorch Implementation of 'BAM-DETR: Boundary-Aligned Moment Detection Transformer for Temporal Sentence Grounding in Videos'☆36Feb 26, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [NeurIPS25] Official Implementation (Pytorch) of "DeepVideo-R1"☆35Feb 22, 2026Updated 4 months ago
- CoFiRec: Coarse-to-Fine Tokenization for Generative Recommendationn☆21Jan 23, 2026Updated 5 months ago
- This is the official code for paper [RePaViT: Scalable Vision Transformer Acceleration via Structural Reparameterization on Feedforward N…☆18Jun 20, 2025Updated last year
- Code for our works: LCSA, C2SLR, and SRM☆23Nov 22, 2024Updated last year
- Project for SNARE benchmark☆11Jun 5, 2024Updated 2 years ago
- Code for "Multi-level Relevance Document Identifier Learning for Generative Retrieval". ACL 2025.☆23Nov 5, 2025Updated 8 months ago
- [ICLR 2025] This repo is the official implementation of our paper "Learning Fine-Grained Representations through Textual Token Disentangl…☆23Jul 28, 2025Updated 11 months ago