☆31Jun 11, 2026Updated 2 months ago
Alternatives and similar repositories for flash-maxsim
Users that are interested in flash-maxsim are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆14Jul 7, 2024Updated 2 years ago
- Fused Triton kernels for late-interaction (MaxSim) scoring☆23Aug 12, 2026Updated 2 weeks ago
- An extensive and commented list of resources on Late-Interaction Multivector Retrieval.☆77Updated this week
- The training codes of Jasper-Token-Compression-600M☆21Nov 19, 2025Updated 9 months ago
- Official repository of TACHIOM.☆63Jul 17, 2026Updated last month
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆21Aug 26, 2025Updated last year
- [SIGIR 2025] The official repo for "Scaling Sparse and Dense Retrieval in Decoder-Only LLMs"☆22Mar 31, 2025Updated last year
- ☆13Jun 2, 2022Updated 4 years ago
- Draft grounded rebuttals to your paper's reviews, with the experiments actually run in your workspace☆17Jul 30, 2026Updated 3 weeks ago
- ModernVBERT is a 250M-parameter vision–language encoder that aligns a text-encoder (Ettin-150M) with a vision-encoder (SigLIP2-B) through…☆16Oct 16, 2025Updated 10 months ago
- Fast search index for SPLADE sparse retrieval models implemented in Python using Numpy and Numba☆39Oct 16, 2025Updated 10 months ago
- bb25 is a fast, self-contained BM25 + Bayesian calibration implementation with a minimal Python API.☆149Mar 17, 2026Updated 5 months ago
- A Rust rewrite of FastKMeans for CPU-based clustering☆17Jun 29, 2026Updated 2 months ago
- Sparse Embedding Compression for Scalable Retrieval in Recommender Systems☆39Nov 21, 2025Updated 9 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Make running benchmark simple yet maintainable, again. Now only supports Korean-based cross-encoder.☆35Aug 21, 2026Updated last week
- Jina VDR is a multilingual, multi-domain benchmark for visual document retrieval☆38Aug 4, 2025Updated last year
- PyTorch implementation of "Sample- and Parameter-Efficient Auto-Regressive Image Models" from CVPR 2025☆14Nov 21, 2025Updated 9 months ago
- Ranking of fine-tuned HF models as base models.☆36Sep 17, 2025Updated 11 months ago
- Tree-based indexes for neural-search☆33Mar 4, 2024Updated 2 years ago
- High-Performance Engine for Multi-Vector Search☆280Updated this week
- MEXMA: Token-level objectives improve sentence representations☆44Jan 6, 2025Updated last year
- A curated list of awesome papers about utilizing large language models for ranking.☆32Apr 12, 2026Updated 4 months ago
- This repository helps you evaluate your models on the FreshStack benchmark!☆34Dec 9, 2025Updated 8 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Model implementation for the contextual embeddings project☆48Jun 2, 2025Updated last year
- NextPlaid, ColGREP: Multi-vector search, from database to coding agents.☆538Updated this week
- [ICML'26] LEMUR reduces multi-vector retrieval for late interaction models such as ColBERT into regular single-vector retrieval.☆33Aug 23, 2026Updated last week
- A modular framework for training and inference of (compressed) multi-vector retrieval across any modality.☆22Apr 4, 2026Updated 4 months ago
- Node-RED nodes to read data from spreadsheet (Excel, ODS, etc.) file☆12Jul 14, 2024Updated 2 years ago
- [ACL 2025] AIR-Bench: Automated Heterogeneous Information Retrieval Benchmark☆167Mar 29, 2026Updated 5 months ago
- Compression for unit-norm embedding vectors using spherical coordinates☆82Jan 23, 2026Updated 7 months ago
- ☆34Feb 27, 2024Updated 2 years ago
- AutoRAG example about benchmarking Korean embeddings.☆46Oct 2, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A massively multilingual modern encoder language model☆152Jan 20, 2026Updated 7 months ago
- ☆21Jun 15, 2026Updated 2 months ago
- ☆13Jan 2, 2022Updated 4 years ago
- 대학생을 위한 IT 스펙 저장소 PRE:FOLIO 클라이언트☆10Jul 19, 2023Updated 3 years ago
- Test-time compute in information retrieval☆59Jul 8, 2025Updated last year
- A fast, streaming-friendly BM25 search engine in Rust with mmap support☆54Mar 19, 2026Updated 5 months ago
- Summarizing Psychology Texts, for Ease of Reference, with LLM / GPT. Attachment, Truama, Polyvagal.☆16Jan 11, 2025Updated last year