[AAAI26]: DS SERVE: The Largest Open Vector Store over Pretain Data; A Framework for Efficient and Scalable Neural Retrieval
☆56Jan 28, 2026Updated 7 months ago
Alternatives and similar repositories for RAG-DS-Serve
Users that are interested in RAG-DS-Serve are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Segmented Code Adjustment Quantization (SAQ)☆27Sep 22, 2025Updated 11 months ago
- ☆35Sep 13, 2025Updated 11 months ago
- [NSDI'26] PolyRL is a reinforcement learning framework for LLM that harvest spot instances on the cloud to reduce cost.☆19Mar 30, 2026Updated 5 months ago
- Memory-Bounded GPU Acceleration for Vector Search☆33Dec 29, 2025Updated 8 months ago
- ☆16Mar 1, 2026Updated 6 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Repository for SIEVE: Effective Filtered Vector Search with Collection of Indexes VLDB '25 submission☆17Jul 30, 2025Updated last year
- ☆40Jun 18, 2025Updated last year
- Vortex: Programmable Sparse Attention for Agents as Algorithm Designers☆68Aug 22, 2026Updated 2 weeks ago
- Query-Adaptive Vector Search☆77Mar 19, 2026Updated 5 months ago
- ☆24Apr 4, 2024Updated 2 years ago
- Block-based Approximate Nearest Neighbor☆35Nov 1, 2021Updated 4 years ago
- ☆23Jun 1, 2025Updated last year
- [SIGMOD2026] Reveal Hidden Pitfalls and Navigate Next Generation of Vector Similarity Search with Task-Centric Benchmarks☆27Dec 31, 2025Updated 8 months ago
- Python package for serving a local search engine. One command to download and serve a datastore---that's it 😎.☆26Jun 6, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A WebUI for Side-by-Side Comparison of Media (Images/Videos) Across Multiple Folders☆26Feb 21, 2025Updated last year
- LM engine is a library for pretraining/finetuning LLMs☆190Updated this week
- ☆16Aug 28, 2025Updated last year
- Building the Virtuous Cycle for AI-driven LLM Systems☆281May 1, 2026Updated 4 months ago
- A cross-modal vector index with fast construction on heterogeneous CPU-GPU environment. Published on DaMoN@SIGMOD 2025.☆16Jul 16, 2025Updated last year
- INCOME: An Easy Repository for Training and Evaluation of Index Compression Methods in Dense Retrieval. Includes BPR and JPQ.☆24Sep 24, 2023Updated 2 years ago
- An LLM inference engine, written in C++☆20Mar 30, 2026Updated 5 months ago
- A toolkit for asynchronously validating dense retriever checkpoints during training.☆27Aug 10, 2023Updated 3 years ago
- Fast full text search for email☆36Feb 15, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A lightweight, high-performance AI/RAG workspace and autonomous agent framework implemented in Go. Inspired by the Leann RAG backend arch…☆36Aug 17, 2026Updated 3 weeks ago
- A benchmark for evaluating LLMs on open-ended CS problems. Exploring the Next Frontier of Computer Science.☆312Aug 21, 2026Updated 2 weeks ago
- XTR/WARP (SIGIR'25) is an extremely fast and accurate retrieval engine based on Stanford's ColBERTv2/PLAID and Google DeepMind's XTR.☆218May 3, 2025Updated last year
- ☆24Oct 23, 2020Updated 5 years ago
- Official repository for "Scaling Retrieval-Based Langauge Models with a Trillion-Token Datastore".☆226Dec 16, 2025Updated 8 months ago
- 🏆 The winner code for Neurips'23 BigANN Competition OOD and Sparse track.☆15Jun 17, 2025Updated last year
- Ada-ef (SIGMOD '26) — Adaptive efSearch for HNSW-based vector search☆20Jun 19, 2026Updated 2 months ago
- How to Train Your Advisor: Steering Black-Box LLMs with Advisor Models☆87Feb 5, 2026Updated 7 months ago
- Repository for the paper: "TiC-LM: A Web-Scale Benchmark for Time-Continual LLM Pretraining" ACL Oral 2025☆24Apr 19, 2026Updated 4 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- How to plot for papers, slides, demos, etc.☆10Apr 7, 2022Updated 4 years ago
- Piecewise Sparse Attention Is Wiser for Efficient Diffusion Transformers☆44Jul 1, 2026Updated 2 months ago
- Official implementation for paper "Navigating Labels and Vectors: A Unified Approach to Filtered Approximate Nearest Neighbor Search"☆38Dec 21, 2024Updated last year
- [SIGMOD'26] Dynamically Detect and Fix Hardness for Efficient Approximate Nearest Neighbor Search☆19Nov 9, 2025Updated 9 months ago
- A High-Throughput Multi-GPU System for Graph-Based Approximate Nearest Neighbor Search☆22Jul 22, 2025Updated last year
- A fast header-only graph-based index for approximate nearest neighbor search (ANNS). https://flatnav.net☆48Jun 4, 2026Updated 3 months ago
- [VLDB 25] Maximum Inner Product is Query-Scaled Nearest Neighbor☆42Oct 31, 2025Updated 10 months ago