Benchmarking library for RAG
☆276Jul 14, 2026Updated last week
Alternatives and similar repositories for bergen
Users that are interested in bergen are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- AutoRAG example about benchmarking Korean embeddings.☆46Oct 2, 2024Updated last year
- Document Ranking with Large Language Models.☆210Feb 14, 2026Updated 5 months ago
- ☆19May 16, 2024Updated 2 years ago
- Official repository of the Seismic library.☆135Jul 6, 2026Updated 2 weeks ago
- SPLADE: sparse neural search (SIGIR21, SIGIR22)☆999May 3, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- RankLLM is a Python toolkit for reproducible information retrieval research using rerankers, with a focus on listwise reranking.☆610Updated this week
- HSEB: Hybrid Search Engine Benchmark☆21Oct 5, 2025Updated 9 months ago
- Sparse Embedding Compression for Scalable Retrieval in Recommender Systems☆39Nov 21, 2025Updated 8 months ago
- Pyserini is a Python toolkit for reproducible information retrieval research with sparse and dense representations.☆2,102Jul 16, 2026Updated last week
- SPRINT Toolkit helps you evaluate diverse neural sparse models easily using a single click on any IR dataset.☆48Jul 25, 2023Updated 3 years ago
- KURE: 고려대학교에서 개발한, 한국어 검색에 특화된 임베딩 모델☆225Apr 14, 2026Updated 3 months ago
- Make running benchmark simple yet maintainable, again. Now only supports Korean-based cross-encoder.☆35Dec 2, 2025Updated 7 months ago
- Inquisitive Parrots for Search☆200Jun 5, 2025Updated last year
- An extensive and commented list of resources on Learned Sparse Retrieval.☆63Updated this week
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Large language models for document ranking.☆75May 20, 2026Updated 2 months ago
- Late Interaction Models Training & Retrieval☆876Updated this week
- Official Code for MIMETIC^2☆13Nov 19, 2024Updated last year
- Korean Sentence Embedding Model Performance Benchmark for RAG☆49Jan 27, 2025Updated last year
- Unified Learned Sparse Retrieval Framework☆68May 13, 2024Updated 2 years ago
- A large-scale multilingual dataset for Information Retrieval. Thorough human-annotations across 18 diverse languages.☆211Jul 31, 2024Updated last year
- Nearly Inference Free Embeddings: make your RAG queries 500x faster☆80Apr 27, 2026Updated 2 months ago
- [ACL 2025] AIR-Bench: Automated Heterogeneous Information Retrieval Benchmark☆167Mar 29, 2026Updated 3 months ago
- [SIGIR'24] Generative Retrieval as Multi-Vector Dense Retrieval☆36Oct 18, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Provides a common interface to many IR ranking datasets.☆390May 28, 2026Updated last month
- Retrieval-Augmented Generation battle!☆66Apr 18, 2026Updated 3 months ago
- SWIM-IR is a Synthetic Wikipedia-based Multilingual Information Retrieval training set with 28 million query-passage pairs spanning 33 la…☆50Nov 13, 2023Updated 2 years ago
- ☆18Jun 16, 2026Updated last month
- Tevatron - Unified Document Retrieval Toolkit across Scale, Language, and Modality. Demo in SIGIR 2023, SIGIR 2025.☆743Jul 18, 2026Updated last week
- MEXMA: Token-level objectives improve sentence representations☆43Jan 6, 2025Updated last year
- Performs benchmarking on two Korean datasets with minimal time and effort.☆45Jan 22, 2026Updated 6 months ago
- ☆63Jan 26, 2025Updated last year
- [SIGIR 2024] The official repo for paper "Planning Ahead in Generative Retrieval: Guiding Autoregressive Generation through Simultaneous …☆32Apr 24, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Starbucks: Improved Training for 2D Matryoshka Embeddings☆25Jun 30, 2025Updated last year
- CLIR version of ColBERT☆73May 28, 2026Updated last month
- The training codes of Jasper-Token-Compression-600M☆20Nov 19, 2025Updated 8 months ago
- ☆14Jul 7, 2024Updated 2 years ago
- Multilingual Dialogue Datasets☆19Aug 18, 2022Updated 3 years ago
- ☆14Jan 10, 2025Updated last year
- A Heterogeneous Benchmark for Information Retrieval. Easy to use, evaluate your models across 15+ diverse IR datasets.☆2,252Oct 16, 2025Updated 9 months ago