Artifact Evaluation for SOSP 2025
☆22Aug 16, 2025Updated last year
Alternatives and similar repositories for HedraRAG_AE
Users that are interested in HedraRAG_AE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A novel system that unifies LLM serving with query optimization to efficiently process batch agentic workflows.☆16Jun 14, 2026Updated 2 months ago
- ☆31Jun 22, 2025Updated last year
- Prefix-Aware Attention for LLM Decoding☆45May 26, 2026Updated 3 months ago
- ☆32Mar 24, 2025Updated last year
- An Open-Source RAG Workload Trace to Optimize RAG Serving Systems☆37Nov 18, 2025Updated 9 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆22Jul 13, 2026Updated last month
- Ada-ef (SIGMOD '26) — Adaptive efSearch for HNSW-based vector search☆19Jun 19, 2026Updated 2 months ago
- ☆51Jul 30, 2025Updated last year
- ☆23Jun 1, 2025Updated last year
- ☆26Apr 13, 2025Updated last year
- A mirror of RLib from lab.☆46Aug 12, 2021Updated 5 years ago
- ☆85Sep 4, 2024Updated last year
- ☆17Nov 11, 2025Updated 9 months ago
- ☆11Aug 9, 2021Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Memory-Bounded GPU Acceleration for Vector Search☆33Dec 29, 2025Updated 8 months ago
- COMP4010 Resources for Spring 2024☆12Jun 5, 2024Updated 2 years ago
- Accelerating Large-Scale Reasoning Model Inference with Sparse Self-Speculative Decoding☆119Dec 2, 2025Updated 8 months ago
- An Optimizing Compiler for Recommendation Model Inference☆26Jun 5, 2025Updated last year
- PipeRAG: Fast Retrieval-Augmented Generation via Algorithm-System Co-design (KDD 2025)☆32Jun 14, 2024Updated 2 years ago
- ☆13Oct 21, 2023Updated 2 years ago
- Artifact for "Apparate: Rethinking Early Exits to Tame Latency-Throughput Tensions in ML Serving" [SOSP '24]☆24Nov 21, 2024Updated last year
- ☆21Jun 9, 2025Updated last year
- ☆21May 11, 2026Updated 3 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆12Dec 1, 2023Updated 2 years ago
- This repository is a collection of awesome things about federated domain generalization, including papers, code, etc.☆17Jul 11, 2026Updated last month
- ☆34Sep 9, 2020Updated 5 years ago
- GPU-accelerated vector query processing system that supports large vector datasets beyond GPU memory.☆41Mar 24, 2024Updated 2 years ago
- Artifact of Chimera☆18May 6, 2025Updated last year
- ☆15Jan 24, 2022Updated 4 years ago
- ☆17Aug 2, 2023Updated 3 years ago
- The collection of papers about Private Evolution☆18Jul 20, 2026Updated last month
- benchmark driver for "Can Learned Models Replace Hash Functions?" VLDB submission☆16Oct 31, 2023Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ROLEX: A Scalable RDMA-oriented Learned Key-Value Store for Disaggregated Memory Systems☆81Jun 8, 2023Updated 3 years ago
- ☆201Jul 15, 2025Updated last year
- [ASPLOS'25] Towards End-to-End Optimization of LLM-based Applications with Ayo☆76Mar 11, 2026Updated 5 months ago
- ☆17Jul 4, 2026Updated last month
- An efficient concurrent graph processing system☆47Oct 27, 2021Updated 4 years ago
- ☆11Sep 4, 2024Updated last year
- [HPCA 2022] GCoD: Graph Convolutional Network Acceleration via Dedicated Algorithm and Accelerator Co-Design☆38Mar 30, 2022Updated 4 years ago