Implementation of the paper Knowledge-Enhanced Dual-stream Zero-shot Composed Image Retrieval (CVPR 2024)
☆20Nov 4, 2024Updated last year
Alternatives and similar repositories for KEDs
Users that are interested in KEDs are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation for "DeltaPhi: Learning Physical Trajectory Residual for PDE Solving"☆13Jun 17, 2024Updated 2 years ago
- ☆10Dec 16, 2023Updated 2 years ago
- The benchmark for "Video Object Segmentation in Panoptic Wild Scenes".☆12Oct 17, 2023Updated 2 years ago
- Visual Delta Generator with Large Multi-modal Model for Semi-supervised Composed Image Retrieval - CVPR2024☆21May 30, 2024Updated 2 years ago
- [ICLR 2024] Official repository for "Vision-by-Language for Training-Free Compositional Image Retrieval"☆89Jul 4, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Evaluation code for "Efficient Emotional Adaptation for Audio-Driven Talking-Head Generation"☆19Mar 10, 2024Updated 2 years ago
- Pytorch implementation of Detective☆13Jul 11, 2024Updated 2 years ago
- [ICCV 2023] - Zero-shot Composed Image Retrieval with Textual Inversion☆198Jul 31, 2025Updated 11 months ago
- Official Pytorch implementation of LinCIR: Language-only Training of Zero-shot Composed Image Retrieval (CVPR 2024)☆148Jan 5, 2026Updated 6 months ago
- Evaluation code and datasets for the ACL 2024 paper, VISTA: Visualized Text Embedding for Universal Multi-Modal Retrieval. The original c…☆48Nov 16, 2024Updated last year
- [ICLR 2026] Official code for paper: TimeSearch-R: Adaptive Temporal Search for Long-Form Video Understanding via Self-Verification Reinf…☆27Jan 29, 2026Updated 5 months ago
- ☆11May 28, 2024Updated 2 years ago
- ☆11Jun 11, 2025Updated last year
- Code for Harnessing the Power of MLLMs for Transferable Text-to-Image Person ReID (CVPR 2024)☆91Jul 13, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [CVPR 2022] Visual Abductive Reasoning☆124Oct 22, 2024Updated last year
- Self-supervised Point Cloud Representation Learning via Separating Mixed Shapes☆21May 23, 2023Updated 3 years ago
- Multi-View prediction enhances GUI Grounding☆21Feb 22, 2026Updated 5 months ago
- The official repository for paper "FlexSelect: Flexible Token Selection for Efficient Long Video Understanding".☆31Sep 19, 2025Updated 10 months ago
- This is the official implementation of "Transferring to Real-World Layouts: A Depth-aware Framework for Scene Adaptation" (Accepted at AC…☆13Aug 24, 2024Updated last year
- Global-to-Local Modeling for Video-based 3D Human Pose and Shape Estimation☆59Jun 21, 2023Updated 3 years ago
- [CVPR2024] CapHuman: Capture Your Moments in Parallel Universes☆99Nov 20, 2024Updated last year
- This repo consists of my implementation of DocFormerV2☆12Mar 31, 2024Updated 2 years ago
- ☆10Dec 14, 2020Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official PyTorch implementation of "Learning with Memory-based Virtual Classes for Deep Metric Learning" (ICCV 2021)☆16Oct 13, 2021Updated 4 years ago
- [NeurIPS 2023] HAP: Structure-Aware Masked Image Modeling for Human-Centric Perception☆44Mar 25, 2024Updated 2 years ago
- ☆197May 9, 2026Updated 2 months ago
- [ACM MM 2022] Patch-based Knowledge Distillation for Lifelong Person Re-Identification☆11Apr 20, 2023Updated 3 years ago
- Implementation (R2R part) for the paper "Iterative Vision-and-Language Navigation"☆18Apr 4, 2024Updated 2 years ago
- Code for SEEG: Semantic Energized Co-speech Gesture Generation☆33Dec 3, 2022Updated 3 years ago
- [BMVC 2023] Zero-shot Composed Text-Image Retrieval☆55Nov 26, 2024Updated last year
- Reward Estimation for Variance Reduction in Deep Reinforcement Learning☆11May 8, 2018Updated 8 years ago
- EVA: Zero-shot Accurate Attributes and Multi-Object Video Editing☆30Mar 29, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Open-source strong baseline for domain generlization re-ID. We will udpate the strong baseline and CFD method~☆10Nov 30, 2021Updated 4 years ago
- [CVPR 2024] Contrasting Intra-Modal and Ranking Cross-Modal Hard Negatives to Enhance Visio-Linguistic Fine-grained Understanding☆56Apr 7, 2025Updated last year
- vue+elementUI 创建的一个好看的UI页面。暂时无js代码,只作为UI展示。☆11Feb 4, 2023Updated 3 years ago
- Code for "Holistic Physics Solver: Learning PDEs in a Unified Spectral-Physical Space"☆25Mar 25, 2026Updated 3 months ago
- [TPAMI 2023] Local-Global Context Aware Transformer for Language-Guided Video Segmentation☆47Jan 20, 2024Updated 2 years ago
- ICCV'2023: Holistic Label Correction for Noisy Multi-Label Classification☆13Oct 29, 2023Updated 2 years ago
- ☆11Dec 30, 2022Updated 3 years ago