Python Implementation of MUVERA (Multi-Vector Retrieval via Fixed Dimensional Encodings)
☆421Dec 10, 2025Updated 8 months ago
Alternatives and similar repositories for muvera-py
Users that are interested in muvera-py are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Generate fixed dimensional embeddings for multi-dimensional vectors in python based on Muvera from Google.☆21Jun 28, 2025Updated last year
- The Python Implementation of CRISP: Clustering Multi-Vector Representations for Denoising and Pruning☆27Jul 27, 2025Updated last year
- ☆15Apr 19, 2026Updated 4 months ago
- [ICML'26] LEMUR reduces multi-vector retrieval for late interaction models such as ColBERT into regular single-vector retrieval.☆33Aug 23, 2026Updated 2 weeks ago
- Make running benchmark simple yet maintainable, again. Now only supports Korean-based cross-encoder.☆36Aug 30, 2026Updated last week
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- bb25 is a fast, self-contained BM25 + Bayesian calibration implementation with a minimal Python API.☆149Mar 17, 2026Updated 5 months ago
- The training codes of Jasper-Token-Compression-600M☆22Nov 19, 2025Updated 9 months ago
- ☆17Aug 13, 2026Updated 3 weeks ago
- Official implementation for paper "Navigating Labels and Vectors: A Unified Approach to Filtered Approximate Nearest Neighbor Search"☆38Dec 21, 2024Updated last year
- ☆16Aug 28, 2025Updated last year
- XTR/WARP (SIGIR'25) is an extremely fast and accurate retrieval engine based on Stanford's ColBERTv2/PLAID and Google DeepMind's XTR.☆219May 3, 2025Updated last year
- A list of multi-vector retrieval resources☆18May 29, 2024Updated 2 years ago
- Late Interaction Models Training & Retrieval☆888Jul 23, 2026Updated last month
- High-Performance Engine for Multi-Vector Search☆281Aug 28, 2026Updated last week
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- An extensive and commented list of resources on Late-Interaction Multivector Retrieval.☆77Aug 29, 2026Updated last week
- HSEB: Hybrid Search Engine Benchmark☆21Oct 5, 2025Updated 11 months ago
- ☆19May 16, 2024Updated 2 years ago
- Official repository of TACHIOM.☆63Updated this week
- PyLate efficient inference engine☆91Jan 7, 2026Updated 8 months ago
- Ada-ef (SIGMOD '26) — Adaptive efSearch for HNSW-based vector search☆20Jun 19, 2026Updated 2 months ago
- AutoRAG example about benchmarking Korean embeddings.☆46Oct 2, 2024Updated last year
- [SIGMOD'26] Dynamically Detect and Fix Hardness for Efficient Approximate Nearest Neighbor Search☆19Nov 9, 2025Updated 9 months ago
- 🎹 Instruct.KR 2025 Summer Meetup: 오픈소스 LLM, vLLM으로 Production까지 🎹☆23Aug 2, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Lightweight Non-Parametric Embedding Fine-Tuning☆42Sep 13, 2025Updated 11 months ago
- [ACL'26 Workshop] KoViDoRe: Korean Visual Document Retrieval Benchmark☆26Jul 2, 2026Updated 2 months ago
- unofficial implementation of MUVERA: Multi-Vector Retrieval via Fixed Dimensional Encodings☆15Feb 18, 2026Updated 6 months ago
- This repository aims to develop CoT Steering based on CoT without Prompting. It focuses on enhancing the model’s latent reasoning capabil…☆117Jun 25, 2025Updated last year
- Pre-train Static Word Embeddings☆111Jun 9, 2026Updated 2 months ago
- Trainable embedding transformation for confidence estimation, feature extraction, explainability and conversion from dense to sparse.☆28Jun 23, 2026Updated 2 months ago
- Fast Multimodal Semantic Deduplication & Filtering☆963May 24, 2026Updated 3 months ago
- Source code for SIGMOD 2020 paper "Improving Approximate Nearest Neighbor Search through Learned Adaptive Early Termination"☆62Jul 17, 2020Updated 6 years ago
- 1-Click is all you need.☆63Apr 29, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Nearly Inference Free Embeddings: make your RAG queries 500x faster☆85Apr 27, 2026Updated 4 months ago
- [SIGMOD2026] Reveal Hidden Pitfalls and Navigate Next Generation of Vector Similarity Search with Task-Centric Benchmarks☆27Dec 31, 2025Updated 8 months ago
- Official repository of the Seismic library.☆136Aug 6, 2026Updated last month
- Train embedding and reranker models for retrieval tasks on Apple Silicon with MLX☆187Sep 18, 2025Updated 11 months ago
- Performs benchmarking on two Korean datasets with minimal time and effort.☆48Aug 6, 2026Updated last month
- Auto Thinking Mode switch for Qwen3 in Open webui☆71May 8, 2025Updated last year
- ☆14Jul 7, 2024Updated 2 years ago