π Modular retrievers for zero-shot multilingual IR.
β30Mar 6, 2024Updated 2 years ago
Alternatives and similar repositories for xm-retrievers
Users that are interested in xm-retrievers are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [SIGIR'24] Generative Retrieval as Multi-Vector Dense Retrievalβ37Oct 18, 2024Updated last year
- β47Mar 27, 2022Updated 4 years ago
- Starbucks: Improved Training for 2D Matryoshka Embeddingsβ25Jun 30, 2025Updated last year
- ACL 2023 Dual-Alignment Pre-training for Cross-lingual Sentence Embeddingβ24Aug 21, 2024Updated 2 years ago
- SWIM-IR is a Synthetic Wikipedia-based Multilingual Information Retrieval training set with 28 million query-passage pairs spanning 33 laβ¦β50Nov 13, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code for "RADCoT: Retrieval-Augmented Distillation to Specialization Models for Generating Chain-of-Thoughts in Query Expansion", LREC-COβ¦β11May 25, 2024Updated 2 years ago
- β10Oct 2, 2024Updated last year
- β11Feb 9, 2024Updated 2 years ago
- YASEM - Yet Another Splade|Sparse Embedder - A simple and efficient library for SPLADE embeddingsβ13May 22, 2025Updated last year
- β16Jun 10, 2024Updated 2 years ago
- β17Jan 5, 2023Updated 3 years ago
- [EMNLP'2024 Findings] Explore generated documents for enhanced IR with LLMs. We enhance BM25 to surpass strong dense retriever on many daβ¦β14Mar 28, 2025Updated last year
- Knowledgeable Embedding: Injecting dynamically updatable entity knowledge into embeddings to enhance RAGβ15Aug 31, 2025Updated last year
- β11Aug 10, 2021Updated 5 years ago
- Managed Kubernetes at scale on DigitalOcean β’ AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Model implementation for the contextual embeddings projectβ49Jun 2, 2025Updated last year
- A multilingual version of MS MARCO passage ranking datasetβ148Oct 19, 2023Updated 2 years ago
- π€ HuggingFace Inference Toolkit for Google Cloud Vertex AI (similar to SageMaker's Inference Toolkit, but for Vertex AI and unofficial)β17Mar 20, 2024Updated 2 years ago
- Implementation of "Efficient Multi-vector Dense Retrieval with Bit Vectors", ECIR 2024β70Oct 21, 2025Updated 11 months ago
- Simple replication of [ColBERT-v1](https://arxiv.org/abs/2004.12832).β83Mar 18, 2024Updated 2 years ago
- A small MNIST-like The Simpsons character database to at least have some fun while training neural networks.β12May 12, 2021Updated 5 years ago
- Source code for paper Grammatical Error Correction in Low-Resource Scenarios (W-NUT 2019)β13Jun 21, 2022Updated 4 years ago
- Cross language information retrieval pipelineβ18Jan 12, 2026Updated 8 months ago
- Code for the ECIR'22 paper "Evaluating the Robustness of Retrieval Pipelines with Query Variation Generators"β18Feb 2, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Soft Prompt Tuning for Augmenting Dense Retrieval with Large Language Modelsβ16Feb 19, 2025Updated last year
- β10May 11, 2024Updated 2 years ago
- A proposed standard `NOCK` for a Parquet format that supports efficient distributed serialization of multiple kinds of graph technologiesβ21Apr 27, 2026Updated 5 months ago
- Prompt Tuning on Graph-augmented Low-resource Text Classification. In TKDE 2024.β15Jan 20, 2025Updated last year
- A simple, easy-to-hack GraphRAG implementationβ15Sep 21, 2024Updated 2 years ago
- Generative Reranker PyTerrierβ18Dec 1, 2025Updated 9 months ago
- The first high-quality, fine-grained error-correction conversation dataset between English second language learner and an educational cβ¦β15Aug 27, 2025Updated last year
- Retrieval-Enhanced Context-Aware Prefix Encoder for Personalized Dialogue Response Generationβ19Aug 26, 2023Updated 3 years ago
- Dual Cross Encoder for Dense Retrievalβ17Mar 15, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- β20Aug 9, 2024Updated 2 years ago
- Hugging Face on Microsoft Azure (documentation, examples and more)β16Updated this week
- GISTEmbed: Guided In-sample Selection of Training Negatives for Text Embeddingsβ47Mar 6, 2024Updated 2 years ago
- PPAT: Progressive Graph Pairwise Attention Network for Event Causality Identificationβ17Jun 7, 2024Updated 2 years ago
- Keyphrase Extraction Prototypesβ15Nov 24, 2016Updated 9 years ago
- Rhythm analysis toolkit in Pythonβ13Jul 5, 2026Updated 2 months ago
- Code and data for: Low Resource Grammatical Error Correction Using Wikipedia Edits (WNUT 2018)β17Jul 16, 2024Updated 2 years ago