[CVPR 2023] VoP: Text-Video Co-operative Prompt Tuning for Cross-Modal Retrieval
☆38Feb 28, 2023Updated 3 years ago
Alternatives and similar repositories for VoP
Users that are interested in VoP are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆16Aug 28, 2024Updated 2 years ago
- Code for paper: Unified Text-to-Image Generation and Retrieval☆15Jul 19, 2026Updated 2 months ago
- Final Drone Repository Ashley Peake and Joe McCalmon WFU 2020☆12Feb 27, 2021Updated 5 years ago
- PyTorch implementation of the AAAI-21 paper "Dual Adversarial Label-aware Graph Neural Networks for Cross-modal Retrieval" and the TPAMI-…☆42Nov 1, 2022Updated 3 years ago
- Pytorch Code for "Unified Coarse-to-Fine Alignment for Video-Text Retrieval" (ICCV 2023)☆66Jun 7, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- This is the implementation for the paper "Generalized Semantic Preserving Hashing for N-Label Cross-Modal Retrieval"☆14Dec 7, 2017Updated 8 years ago
- Video embeddings for retrieval with natural language queries☆342Feb 15, 2023Updated 3 years ago
- A curated list of deep learning resources for video-text retrieval.☆645Oct 20, 2023Updated 2 years ago
- Label Embedding Online Hashing for Cross-Modal Retrieval☆13Sep 22, 2025Updated last year
- ☆10Nov 23, 2023Updated 2 years ago
- The source code of "Bit-aware Semantic Transformer Hashing for Multi-modal Retrieval." (Accepted by SIGIR 2022)☆18Sep 15, 2022Updated 4 years ago
- On-Device Domain Generalization☆48Nov 9, 2022Updated 3 years ago
- This repository contains the author's implementation in PyTorch for the paper "Adaptive Label-aware Graph Convolutional Networks for Cros…☆15Dec 6, 2021Updated 4 years ago
- [EMNLP 2024 Main] MaPPER: Multimodal Prior-guided Parameter Efficient Tuning for Referring Expression Comprehension☆16Jan 6, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- BiC-Net: Learning Efficient Spatio-Temporal Relation for Text-Video Retrieval☆28Jul 22, 2022Updated 4 years ago
- The Demo of Our CVPR paper "Cross-Modality Binary Code Learning via Fusion Similarity Hashing"☆14Sep 7, 2017Updated 9 years ago
- 😎 All your need for future is FollowGPT.☆13Nov 8, 2023Updated 2 years ago
- An official implementation for "CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval"☆1,033Apr 12, 2024Updated 2 years ago
- [ICCV2023] Tem-adapter: Adapting Image-Text Pretraining for Video Question Answer☆37Oct 18, 2023Updated 2 years ago
- This repository contains the official implementation of CoMix (NeurIPS 2021) https://arxiv.org/pdf/2110.15128.pdf.☆22Jan 12, 2022Updated 4 years ago
- Evaluate robustness of adaptation methods on large vision-language models☆19Aug 23, 2023Updated 3 years ago
- Code and benchmarks for the Semantic Video Retrieval Task☆53Oct 18, 2022Updated 3 years ago
- Official Implementation of Paper FOLDER (ICCV2025) and Turbo (ECCV2024)☆15Jun 27, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Ask&Confirm: Active Detail Enriching for Cross-Modal Retrieval with Partial Query (ICCV2021)☆20Dec 4, 2021Updated 4 years ago
- Awesome paper for multi-modal llm with grounding ability☆21Oct 11, 2025Updated 11 months ago
- The source code for the CVPR2020 paper "Creating Something from Nothing: Unsupervised Knowledge Distillation for Cross-Modal Hashing".☆24Oct 10, 2020Updated 5 years ago
- [ICME 2024 Oral] DARA: Domain- and Relation-aware Adapters Make Parameter-efficient Tuning for Visual Grounding☆22Feb 26, 2025Updated last year
- Benchmark data for "Rethinking Benchmarks for Cross-modal Image-text Retrieval" (SIGIR 2023)☆28Apr 24, 2023Updated 3 years ago
- ☆26Jan 12, 2022Updated 4 years ago
- ☆82Nov 6, 2023Updated 2 years ago
- Official implementation of PCS in essay "Prompt Vision Transformer for Domain Generalization"☆49Jan 29, 2023Updated 3 years ago
- Cross Modal Retrieval with Querybank Normalisation☆57Nov 21, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The code for the paper "Hybrid Contrastive Quantization for Efficient Cross-View Video Retrieval" (WWW'22, Oral).☆17Mar 8, 2022Updated 4 years ago
- [CVPR 2023] Official repository of paper titled "Fine-tuned CLIP models are efficient video learners".☆310Apr 3, 2024Updated 2 years ago
- ☆13Oct 4, 2023Updated 2 years ago
- COLA: Evaluate how well your vision-language model can Compose Objects Localized with Attributes!☆25May 14, 2026Updated 4 months ago
- ☆12Aug 5, 2022Updated 4 years ago
- [ICCV 2023 & AAAI 2023] Binary Adapters & FacT, [Tech report] Convpass☆202Aug 1, 2023Updated 3 years ago
- ☆259Dec 10, 2022Updated 3 years ago