☆19Aug 16, 2025Updated last year
Alternatives and similar repositories for Jenga-SOSP25-AE
Users that are interested in Jenga-SOSP25-AE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Artifact for "Marconi: Prefix Caching for the Era of Hybrid LLMs" [MLSys '25 Outstanding Paper Award, Honorable Mention]☆70Mar 5, 2025Updated last year
- [PACT'24] GraNNDis. A fast and unified distributed graph neural network (GNN) training framework for both full-batch (full-graph) and min…☆10Aug 13, 2024Updated 2 years ago
- ☆24Mar 7, 2025Updated last year
- Official implementation of "SentenceKV: Efficient LLM Inference via Sentence-Level Semantic KV Caching" (COLM 2025). A novel KV cache com…☆15Sep 29, 2025Updated last year
- Prototyp MegaScale-Infer: Serving Mixture-of-Experts at Scale with Disaggregated Expert Parallelism☆34Apr 4, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Preview Code for Continuum Paper☆108Aug 13, 2026Updated last month
- Implementation of "RaanA: A Fast, Flexible, and Data-Efficient Post-Training Quantization Algorithm"☆18Apr 11, 2025Updated last year
- ☆22Jul 13, 2026Updated 2 months ago
- PPoPP24 AGAThA: Fast and Efficient GPU Acceleration of Guided Sequence Alignment for Long Read Mapping☆22May 8, 2024Updated 2 years ago
- A High-Throughput Multi-GPU System for Graph-Based Approximate Nearest Neighbor Search☆21Jul 22, 2025Updated last year
- ☆26Jan 21, 2026Updated 8 months ago
- Sharing the codebase and steps for artifact evaluation for ISCA 2023 paper☆16Feb 20, 2024Updated 2 years ago
- TaskMet Task-driven Metric Learning for Model Learning☆21Feb 9, 2024Updated 2 years ago
- [ICML 2025] Efficiently Serving Large Multimodal Models Using EPD Disaggregation☆28Jul 11, 2026Updated 2 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆43Oct 12, 2025Updated 11 months ago
- ☆14Nov 11, 2025Updated 10 months ago
- The official code implementation for paper "PM-KVQ: Progressive Mixed-precision KV Cache Quantization for Long-CoT LLMs"☆30May 24, 2025Updated last year
- An open platform for exploring scale-up network systems.☆23Mar 16, 2026Updated 6 months ago
- ☆30Sep 11, 2026Updated 3 weeks ago
- A paper review list for computer architecture and systems research, maintained by the LEMONADE group at Peking University.☆22Sep 28, 2026Updated last week
- Performant kernels for symmetric tensors☆17Aug 22, 2024Updated 2 years ago
- ☆12Sep 18, 2024Updated 2 years ago
- Elana: A Simple Energy & Latency Analyzer for LLMs☆16Apr 3, 2026Updated 6 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆34Sep 9, 2020Updated 6 years ago
- ☆23Jul 7, 2026Updated 3 months ago
- 计算机体系结构:量化研究方法(中文版)附录D-M☆23Aug 22, 2023Updated 3 years ago
- ☆10Oct 8, 2021Updated 5 years ago
- RisGraph: A Real-Time Streaming System for Evolving Graphs to Support Sub-millisecond Per-update Analysis at Millions Ops/s☆39May 11, 2022Updated 4 years ago
- ☆18May 3, 2024Updated 2 years ago
- A docker image for One Student One Chip's debug exam☆10Sep 22, 2023Updated 3 years ago
- p2p db functions for chat☆13May 14, 2018Updated 8 years ago
- Cluster Far Mem, framework to execute single job and multi job experiments using fastswap☆21Jan 12, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Artifact of Chimera☆18May 6, 2025Updated last year
- a student trainning project for HLS and transformer☆11Oct 19, 2022Updated 3 years ago
- Sometimes all you need is a single file.☆15Sep 29, 2020Updated 6 years ago
- ☆14Jun 22, 2022Updated 4 years ago
- The official implementation for the intra-stage fusion technique introduced in https://arxiv.org/abs/2409.13221☆32Apr 22, 2025Updated last year
- A demo of using automerge to build a chat client.☆16Aug 15, 2022Updated 4 years ago
- DHCP Client in Rust☆11Apr 10, 2023Updated 3 years ago