[NeurIPS 2023] The official implementation of paper "Prototype-based Aleatoric Uncertainty Quantification for Cross-modal Retrieval" accepted by NeurIPS' 2023.
β28May 14, 2024Updated 2 years ago
Alternatives and similar repositories for PAU
Users that are interested in PAU are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV'23] UATVR: Uncertainty-Adaptive Text-Video Retrievalβ13Nov 5, 2023Updated 2 years ago
- π¦Ύ SeClaw: The Security Armored Personal AI Assistantβ31Mar 18, 2026Updated 4 months ago
- The code of "Image-text Retrieval via Preserving Main Semantic of Vision" in ICME 2023.β15Dec 25, 2023Updated 2 years ago
- Official Pytorch implementation of "Improved Probabilistic Image-Text Representations" (ICLR 2024)β63May 26, 2024Updated 2 years ago
- Pytorch Code for "Unified Coarse-to-Fine Alignment for Video-Text Retrieval" (ICCV 2023)β66Jun 7, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Text Proxy: Decomposing Retrieval from a 1-to-N Relationship into N 1-to-1 Relationships for Text-Video Retrieval -- AAAI2025β21May 8, 2026Updated 3 months ago
- [DMLR 2024] Benchmarking Robustness of Multimodal Image-Text Models under Distribution Shiftβ39Jan 25, 2024Updated 2 years ago
- Repository of "Improving Cross-Modal Retrieval With Set of Diverse Embeddings" (CVPR'23, Highlight)β41Nov 15, 2023Updated 2 years ago
- [NeurIPS 2025] The official implementation of the paper "DRIFT: Dynamic Rule-Based Defense with Injection Isolation for Securing LLM Agenβ¦β59Jul 16, 2026Updated 3 weeks ago
- [CVPR 2023] Enlarge Instance-specific and Class-specific Information for Open-set Action Recognitionβ31Apr 19, 2023Updated 3 years ago
- The code of the paper of "A Differentiable Semantic Metric Approximation in Probabilistic Embedding for Cross-Modal Retrieval" accepted bβ¦β19Jan 16, 2024Updated 2 years ago
- [CVPR 2025] DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrievalβ22Jun 23, 2025Updated last year
- [NeurIPS 2023] Bootstrapping Vision-Language Learning with Decoupled Language Pre-trainingβ26Dec 5, 2023Updated 2 years ago
- β30Aug 14, 2023Updated 2 years ago
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [ACL 2025] The official implementation of the paper "PIGuard: Prompt Injection Guardrail via Mitigating Overdefense for Free".β81Dec 4, 2025Updated 8 months ago
- offical implementation of "Calibrating Multimodal Learning" on ICML 2023β20Jun 5, 2023Updated 3 years ago
- The code of the paper "Negative Pre-aware for Noisy Cross-modal Matching" in AAAI 2024.β31Jul 22, 2026Updated 3 weeks ago
- β19Jul 28, 2025Updated last year
- β82Nov 6, 2023Updated 2 years ago
- Official github repo for ICCV2023 paper 'Multi-event Video-Text Retrieval'β20Feb 16, 2024Updated 2 years ago
- Code implementation of paper "MUSE: Mamba is Efficient Multi-scale Learner for Text-video Retrieval (AAAI2025)"β26Feb 2, 2025Updated last year
- [ICLR 2024] Official repository for "Vision-by-Language for Training-Free Compositional Image Retrieval"β89Jul 4, 2024Updated 2 years ago
- [CVPR 2024] TeachCLIP for Text-to-Video Retrievalβ42May 7, 2025Updated last year
- Open source password manager - Proton Pass β’ AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- [AAAI 2024] GMMFormer: Gaussian-Mixture-Model Based Transformer for Efficient Partially Relevant Video Retrievalβ21May 10, 2024Updated 2 years ago
- ProbVLM: Probabilistic Adapter for Frozen Vision-Language Modelsβ47Dec 21, 2023Updated 2 years ago
- Official PyTorch implementation of our TGRS paper: Deep Adaptive Pansharpening via Uncertainty-aware Image Fusion.β14Aug 7, 2023Updated 3 years ago
- [IJCAI 2023] Text-Video Retrieval with Disentangled Conceptualization and Set-to-Set Alignmentβ53Apr 9, 2024Updated 2 years ago
- This is a summary of research on noisy correspondence. There may be omissions. If anything is missing please get in touch with us. Our emβ¦β85May 24, 2026Updated 2 months ago
- β10Nov 23, 2023Updated 2 years ago
- [ICLR 2025] This repo is the official implementation of our paper "Learning Fine-Grained Representations through Textual Token Disentanglβ¦β23Jul 28, 2025Updated last year
- [ICCV2023] Tem-adapter: Adapting Image-Text Pretraining for Video Question Answerβ37Oct 18, 2023Updated 2 years ago
- Reason-before-Retrieve: One-Stage Reflective Chain-of-Thoughts for Training-Free Zero-Shot Composed Image Retrieval [CVPR 2025 Highlight]β72Jul 8, 2025Updated last year
- End-to-end encrypted cloud storage - Proton Drive β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- [WACV 2024] Enhancing Multimodal Compositional Reasoning of Visual Language Models with Generative Negative Mining, WACV 2024β13Jan 3, 2024Updated 2 years ago
- [CVPR 2024] Contrasting Intra-Modal and Ranking Cross-Modal Hard Negatives to Enhance Visio-Linguistic Fine-grained Understandingβ56Apr 7, 2025Updated last year
- [ICLR 2025] TempMe: Video Temporal Token Merging for Efficient Text-Video Retrievalβ27Feb 13, 2025Updated last year
- [CVPR2024] The code of "UniPT: Universal Parallel Tuning for Transfer Learning with Efficient Parameter and Memory"β71Oct 15, 2024Updated last year
- Implementation of "DIME-FM: DIstilling Multimodal and Efficient Foundation Models"β15Oct 12, 2023Updated 2 years ago
- Bidirectional Likelihood Estimation with Multi-Modal Large Language Models for Text-Video Retrieval (ICCV 2025 Highlight)β27Aug 1, 2025Updated last year
- Official PyTorch implementation of the paper "CoVR: Learning Composed Video Retrieval from Web Video Captions".β119Apr 21, 2026Updated 3 months ago