Official repository of "Chatting Makes Perfect: Chat-based Image Retrieval"
☆32Feb 5, 2025Updated last year
Alternatives and similar repositories for ChatIR
Users that are interested in ChatIR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV 2023] Simple Baselines for Interactive Video Retrieval with Questions and Answers☆20Apr 16, 2024Updated 2 years ago
- [CVPR 2025] LamRA: Large Multimodal Model as Your Advanced Retrieval Assistant☆182Jul 7, 2025Updated last year
- [NeurIPS 2019] Drill-down: Interactive Retrieval of Complex Scenes using Natural Language Queries☆12Apr 15, 2022Updated 4 years ago
- Pytorch Implementation of LLaVA-ReID: Selective Multi-image Questioner for Interactive Person Re-Identification☆107Nov 20, 2025Updated 8 months ago
- This repository is the official implementation of Dataset Condensation with Contrastive Signals (DCC), accepted at ICML 2022.☆22Jun 8, 2022Updated 4 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code for our paper: "Where's Waldo: Diffusion Features For Personalized Segmentation and Retrieval".☆14Feb 26, 2025Updated last year
- Visualize KITTI360 sequences on ROS with full tf support.☆10Apr 21, 2023Updated 3 years ago
- Look, Compare, Decide: Alleviating Hallucination in Large Vision-Language Models via Multi-View Multi-Path Reasoning☆24Sep 9, 2024Updated last year
- a simple questionnaire Flask web app☆19May 1, 2023Updated 3 years ago
- Curated List of NLP tutorials☆30Feb 27, 2025Updated last year
- Collect data sets and research papers in the field of 3D computer vision tasks with implemented repositories.☆23Jun 26, 2020Updated 6 years ago
- [ACM MM 2024] Improving Composed Image Retrieval via Contrastive Learning with Scaling Positives and Negatives☆39Sep 9, 2025Updated 10 months ago
- ☆12Feb 2, 2024Updated 2 years ago
- This is a summary of research on noisy correspondence. There may be omissions. If anything is missing please get in touch with us. Our em…☆86May 24, 2026Updated last month
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Human-centered Interactive Learning via MLLMs for Text-to-Image Person Re-identification (CVPR 2025 Pytorch Code)☆49Jul 19, 2025Updated last year
- Code for Modeling Thousands of Human Annotators for Generalizable Text-to-Image Person Re-identification (CVPR2025)☆50Nov 4, 2025Updated 8 months ago
- ☆12May 20, 2019Updated 7 years ago
- ☆11Nov 28, 2022Updated 3 years ago
- [NeurIPS 2023 D&B] VidChapters-7M: Video Chapters at Scale☆211Nov 13, 2023Updated 2 years ago
- COLA: Evaluate how well your vision-language model can Compose Objects Localized with Attributes!☆25May 14, 2026Updated 2 months ago
- A lightweight open-source package to fine-tune embedding models.☆22Feb 4, 2024Updated 2 years ago
- [EMNLP 2024 Industry track] MERLIN : Multimodal Embedding Refinement via LLM-based Iterative Navigation for Text-Video Retrieval-Rerank P…☆14Mar 4, 2025Updated last year
- ☆17Jan 30, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Context-I2W: Mapping Images to Context-dependent words for Accurate Zero-Shot Composed Image Retrieval [AAAI 2024 Oral]☆54May 27, 2025Updated last year
- Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency☆62Jun 6, 2025Updated last year
- ☆25May 13, 2024Updated 2 years ago
- Mixture-of-Embeddings-Experts☆122Jul 21, 2020Updated 6 years ago
- A list of multi-vector retrieval resources☆19May 29, 2024Updated 2 years ago
- Multi-Modal Mutual Information (MuMMI) Training for Robust Self-Supervised Deep Reinforcement Learning☆13Jun 28, 2022Updated 4 years ago
- The implementation of FINER-MLLM, which is accepted by MM2024.☆18Oct 8, 2024Updated last year
- [NeurIPS 2023] The official implementation of paper "Prototype-based Aleatoric Uncertainty Quantification for Cross-modal Retrieval" acce…☆28May 14, 2024Updated 2 years ago
- ☆18Feb 20, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Codes for ICLR 2025 Paper: Towards Semantic Equivalence of Tokenization in Multimodal LLM☆81Apr 19, 2025Updated last year
- Corona Virus Data Visuzalization Platform☆17May 15, 2021Updated 5 years ago
- [CVPR 2026] Pytorch Code for the paper "Bootstrapping Multi-view Learning for Test-time Noisy Correspondence"☆15Jul 1, 2026Updated 3 weeks ago
- ☆15Jun 22, 2022Updated 4 years ago
- Python GUI application that generates images based on user prompts using the StableDiffusionPipeline model from the diffusers module. The…☆14May 28, 2023Updated 3 years ago
- [ICCV 2025] This repo is the official implementation of "Multi-Object Sketch Animation by Scene Decomposition and Motion Planning"☆28Jul 30, 2025Updated 11 months ago
- ☆10Nov 6, 2018Updated 7 years ago