Bayesian probability transforms for BM25 retrieval scores
☆77Jun 20, 2026Updated 3 months ago
Alternatives and similar repositories for bayesian-bm25
Users that are interested in bayesian-bm25 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- No code tool for finetuning embedding models☆35Updated this week
- HSEB: Hybrid Search Engine Benchmark☆21Oct 5, 2025Updated 11 months ago
- A small, crazy fast hybrid search engine written in Rust.☆34Jun 27, 2026Updated 2 months ago
- High-Performance Engine for Multi-Vector Search☆282Sep 10, 2026Updated last week
- User Behavior Insights standard schema specification☆47Oct 8, 2025Updated 11 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A Python library for creating adversarial splits☆14Jul 24, 2022Updated 4 years ago
- Improve your OpenSearch, Elasticsearch, Solr, Vectara, Algolia and Custom Search search quality.☆342Updated this week
- C++ inference wrappers for running blazing fast embedding services on your favourite serverless like AWS Lambda. By Prithivi Da, PRs welc…☆24Mar 4, 2024Updated 2 years ago
- A framework for benchmarking embedding models in hybrid search scenarios (BM25 + vector search) using Weaviate.☆40Aug 19, 2026Updated last month
- PyLate efficient inference engine☆91Jan 7, 2026Updated 8 months ago
- Late Interaction Models Training & Retrieval☆895Jul 23, 2026Updated last month
- Code and Data of the paper: "Redefining Retrieval Evaluation in the Era of LLMs"☆16Oct 27, 2025Updated 10 months ago
- Nearly Inference Free Embeddings: make your RAG queries 500x faster☆86Apr 27, 2026Updated 4 months ago
- Cheat at search with LLMs course materials☆51Sep 14, 2026Updated last week
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- IO-aware batched K-Means for Apple Silicon, ported from Flash-KMeans (Triton/CUDA) to pure MLX. Up to 94x faster than sklearn.☆18Mar 22, 2026Updated 6 months ago
- Fast Diversification for Search & Retrieval☆499May 24, 2026Updated 3 months ago
- Fetches transcripts from YouTube videos, including private ones with granted access, and optionally downloads the videos. Does not suppor…☆18Apr 17, 2024Updated 2 years ago
- [ICML'26] LEMUR reduces multi-vector retrieval for late interaction models such as ColBERT into regular single-vector retrieval.☆33Sep 13, 2026Updated last week
- The Python Implementation of CRISP: Clustering Multi-Vector Representations for Denoising and Pruning☆27Jul 27, 2025Updated last year
- Full text search that feels like a numpy array☆312May 4, 2026Updated 4 months ago
- SPRINT Toolkit helps you evaluate diverse neural sparse models easily using a single click on any IR dataset.☆48Jul 25, 2023Updated 3 years ago
- Neural Solr = Solr 9 + Mighty Inference + Node☆18Jun 9, 2022Updated 4 years ago
- Optimize and Enhance Your Search Quality☆16May 1, 2026Updated 4 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Mini Callcenter Simulator simulates a call center and takes into account many parameters not covered by the Erlang C formula.☆13Jul 25, 2026Updated last month
- Fast BM25 search in Python, powered by Numpy and Numba☆1,791Updated this week
- MEXMA: Token-level objectives improve sentence representations☆44Jan 6, 2025Updated last year
- Towards an open source stack for e-commerce search☆155Mar 21, 2026Updated 6 months ago
- State-of-the-art paired encoder and decoder models (17M-1B params)☆79Aug 6, 2025Updated last year
- XTR: Rethinking the Role of Token Retrieval in Multi-Vector Retrieval☆65Jun 20, 2024Updated 2 years ago
- Contextualized per-token embeddings☆41Sep 2, 2026Updated 3 weeks ago
- Sparse Embedding Compression for Scalable Retrieval in Recommender Systems☆39Nov 21, 2025Updated 10 months ago
- ⚡️A Blazing-Fast Python Library for Ranking Evaluation, Comparison, and Fusion 🐍☆701Aug 7, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Build & Install custom C/C++ extension package for Pyodide☆23Aug 13, 2022Updated 4 years ago
- allowing R users to work with dlib through Rcpp☆13Apr 11, 2018Updated 8 years ago
- ☆42Jan 29, 2026Updated 7 months ago
- Parallel wasm Barnes-Hut t-SNE implementation written in Rust.☆23Aug 6, 2026Updated last month
- Model implementation for the contextual embeddings project☆49Jun 2, 2025Updated last year
- Train embedding and reranker models for retrieval tasks on Apple Silicon with MLX☆187Sep 18, 2025Updated last year
- [SIGIR 2025] The official repo for "Scaling Sparse and Dense Retrieval in Decoder-Only LLMs"☆22Mar 31, 2025Updated last year