Bayesian probability transforms for BM25 retrieval scores
☆77Jun 20, 2026Updated 2 months ago
Alternatives and similar repositories for bayesian-bm25
Users that are interested in bayesian-bm25 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- No code tool for finetuning embedding models☆32Aug 6, 2026Updated 3 weeks ago
- HSEB: Hybrid Search Engine Benchmark☆21Oct 5, 2025Updated 10 months ago
- A small, crazy fast hybrid search engine written in Rust.☆29Jun 27, 2026Updated 2 months ago
- High-Performance Engine for Multi-Vector Search☆281Updated this week
- A Python library for creating adversarial splits☆14Jul 24, 2022Updated 4 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Improve your OpenSearch, Elasticsearch, Solr, Vectara, Algolia and Custom Search search quality.☆342Updated this week
- C++ inference wrappers for running blazing fast embedding services on your favourite serverless like AWS Lambda. By Prithivi Da, PRs welc…☆24Mar 4, 2024Updated 2 years ago
- ☆15Apr 14, 2023Updated 3 years ago
- A framework for benchmarking embedding models in hybrid search scenarios (BM25 + vector search) using Weaviate.☆40Aug 19, 2026Updated last week
- Late Interaction Models Training & Retrieval☆888Jul 23, 2026Updated last month
- Code and Data of the paper: "Redefining Retrieval Evaluation in the Era of LLMs"☆16Oct 27, 2025Updated 10 months ago
- Nearly Inference Free Embeddings: make your RAG queries 500x faster☆85Apr 27, 2026Updated 4 months ago
- IO-aware batched K-Means for Apple Silicon, ported from Flash-KMeans (Triton/CUDA) to pure MLX. Up to 94x faster than sklearn.☆17Mar 22, 2026Updated 5 months ago
- Fast Diversification for Search & Retrieval☆497May 24, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Make running benchmark simple yet maintainable, again. Now only supports Korean-based cross-encoder.☆36Updated this week
- [ICML'26] LEMUR reduces multi-vector retrieval for late interaction models such as ColBERT into regular single-vector retrieval.☆33Aug 23, 2026Updated last week
- Pre-train Static Word Embeddings☆111Jun 9, 2026Updated 2 months ago
- The Python Implementation of CRISP: Clustering Multi-Vector Representations for Denoising and Pruning☆27Jul 27, 2025Updated last year
- Full text search that feels like a numpy array☆312May 4, 2026Updated 3 months ago
- SPRINT Toolkit helps you evaluate diverse neural sparse models easily using a single click on any IR dataset.☆48Jul 25, 2023Updated 3 years ago
- Neural Solr = Solr 9 + Mighty Inference + Node☆18Jun 9, 2022Updated 4 years ago
- ☆26Feb 11, 2025Updated last year
- Fast BM25 search in Python, powered by Numpy and Numba☆1,778Aug 25, 2026Updated last week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Improving Text Embedding of Language Models Using Contrastive Fine-tuning☆64Aug 2, 2024Updated 2 years ago
- Set-Encoder: Permutation-Invariant Inter-Passage Attention for Listwise Passage Re-Ranking with Cross-Encoders☆19May 23, 2025Updated last year
- MEXMA: Token-level objectives improve sentence representations☆44Jan 6, 2025Updated last year
- Powerful unsupervised domain adaptation method for dense retrieval. Requires only unlabeled corpus and yields massive improvement: "GPL: …☆342Jul 6, 2023Updated 3 years ago
- Towards an open source stack for e-commerce search☆154Mar 21, 2026Updated 5 months ago
- State-of-the-art paired encoder and decoder models (17M-1B params)☆78Aug 6, 2025Updated last year
- XTR: Rethinking the Role of Token Retrieval in Multi-Vector Retrieval☆64Jun 20, 2024Updated 2 years ago
- Contextualized per-token embeddings☆41Aug 22, 2026Updated last week
- An extensive and commented list of resources on Learned Sparse Retrieval.☆64Aug 4, 2026Updated 3 weeks ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ⚡️A Blazing-Fast Python Library for Ranking Evaluation, Comparison, and Fusion 🐍☆694Aug 7, 2025Updated last year
- Build & Install custom C/C++ extension package for Pyodide☆23Aug 13, 2022Updated 4 years ago
- Coord: A Unified Interface for All Models☆18Jun 19, 2026Updated 2 months ago
- Parallel wasm Barnes-Hut t-SNE implementation written in Rust.☆23Aug 6, 2026Updated 3 weeks ago
- Model implementation for the contextual embeddings project☆48Jun 2, 2025Updated last year
- Train embedding and reranker models for retrieval tasks on Apple Silicon with MLX☆187Sep 18, 2025Updated 11 months ago
- [SIGIR 2025] The official repo for "Scaling Sparse and Dense Retrieval in Decoder-Only LLMs"☆22Mar 31, 2025Updated last year