[ACL 2025 Oral] π₯π₯ MegaPairs: Massive Data Synthesis for Universal Multimodal Retrieval
β250Nov 6, 2025Updated 10 months ago
Alternatives and similar repositories for MegaPairs
Users that are interested in MegaPairs are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repo contains the code for "VLM2Vec / MMEB" [ICLR 2025], "VLM2Vec-V2 / MMEB-V2" [TMLR 2026], and "MMEB-V3" [COLM 2026]β686Updated this week
- LLaVE: Large Language and Vision Embedding Models with Hardness-Weighted Contrastive Learningβ79May 23, 2025Updated last year
- β26Jun 22, 2026Updated 3 months ago
- official code for "Modality Curation: Building Universal Embeddings for Advanced Multimodal Information Retrieval"β42Jul 4, 2025Updated last year
- Official implement of CIKM2025: γUniECS: Unified Multimodal E-Commerce Search Framework with Gated Cross-modal Fusionγβ21Sep 17, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Not a neutral survey β a field manual for engineers who build, train, and ship multimodal retrieval at production scale. The C-L-I triangβ¦β83Apr 20, 2026Updated 5 months ago
- E5-V: Universal Embeddings with Multimodal Large Language Modelsβ274Dec 10, 2025Updated 9 months ago
- [ACM MM 2025] The official code of "Breaking the Modality Barrier: Universal Embedding Learning with Multimodal LLMs"β105Dec 8, 2025Updated 9 months ago
- [CVPR 2025] LamRA: Large Multimodal Model as Your Advanced Retrieval Assistantβ183Jul 7, 2025Updated last year
- Evaluation code and datasets for the ACL 2024 paper, VISTA: Visualized Text Embedding for Universal Multi-Modal Retrieval. The original cβ¦β48Nov 16, 2024Updated last year
- Reason-before-Retrieve: One-Stage Reflective Chain-of-Thoughts for Training-Free Zero-Shot Composed Image Retrieval [CVPR 2025 Highlight]β73Jul 8, 2025Updated last year
- Official Repo: AutoResearchBench: Benchmarking AI Agents on Complex Scientific Literature Discoveryβ71Apr 24, 2026Updated 5 months ago
- β45Jan 12, 2026Updated 8 months ago
- β154Nov 17, 2025Updated 10 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- β59Feb 27, 2025Updated last year
- Collection of Composed Image Retrieval (CIR) papers.β373Updated this week
- ABC: Achieving Better Control of Multimodal Embeddings using VLMs [TMLR2025]β20Aug 21, 2025Updated last year
- [AAAI 2026 Oral] The official code of "UniME-V2: MLLM-as-a-Judge for Universal Multimodal Embedding Learning"β76Dec 8, 2025Updated 9 months ago
- Official code for paper "UniIR: Training and Benchmarking Universal Multimodal Information Retrievers" (ECCV 2024)β186Oct 1, 2024Updated last year
- Data Synthesis for Deep Research Based on Semi-Structured Dataβ217Jul 14, 2026Updated 2 months ago
- β24Jul 23, 2025Updated last year
- A comprehensive survey of Composed Multi-modal Retrieval (CMR), including Composed Image Retrieval (CIR) and Composed Video Retrieval (CVβ¦β92Jul 31, 2026Updated last month
- β13Nov 26, 2021Updated 4 years ago
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [WWW 2025 Oral] ImageScope: Unifying Language-Guided Image Retrieval via Large Multimodal Model Collective Reasoningβ21Jul 2, 2025Updated last year
- Executive Memory for Coherent Long-Horizon Reasoning!β87Jan 14, 2026Updated 8 months ago
- π₯π₯First-ever hour scale video understanding modelsβ630Jul 14, 2025Updated last year
- β20Mar 5, 2025Updated last year
- [ICLR 2025 Spotlight] OmniCorpus: A Unified Multimodal Corpus of 10 Billion-Level Images Interleaved with Textβ429May 5, 2025Updated last year
- Retrieval and Retrieval-augmented LLMsβ12,191Aug 24, 2026Updated last month
- Code for "CAFe: Unifying Representation and Generation with Contrastive-Autoregressive Finetuning"β34Mar 26, 2025Updated last year
- New generation of CLIP with strong fine grained discrimination capability, ICML2026 and ICML2025β735Jun 16, 2026Updated 3 months ago
- [ICML'24 Oral] "MagicLens: Self-Supervised Image Retrieval with Open-Ended Instructions"β212Oct 28, 2024Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer β’ AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Repo for WWW 2022 paper: Progressively Optimized Bi-Granular Document Representation for Scalable Embedding Based Retrievalβ16Mar 1, 2022Updated 4 years ago
- [ICLR 2026] EditScore: Unlocking Online RL for Image Editing via High-Fidelity Reward Modelingβ258Mar 20, 2026Updated 6 months ago
- β2,037Sep 30, 2025Updated 11 months ago
- [NeurIPS 2025] Scaling Language-centric Omnimodal Representation Learningβ49Apr 13, 2026Updated 5 months ago
- β1,387Jun 23, 2026Updated 3 months ago
- β69Aug 14, 2025Updated last year
- [NeurIPS 2023] HAP: Structure-Aware Masked Image Modeling for Human-Centric Perceptionβ44Mar 25, 2024Updated 2 years ago