Accelerating AI Training and Inference from Storage Perspective (Must-read Papers on Storage for AI)
☆67Sep 9, 2026Updated 2 weeks ago
Alternatives and similar repositories for Storage-for-AI-Paper
Users that are interested in Storage-for-AI-Paper are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The code based on vLLM for the paper “ Cost-Efficient Large Language Model Serving for Multi-turn Conversations with CachedAttention”.☆11Sep 19, 2024Updated 2 years ago
- ☆15Jan 21, 2023Updated 3 years ago
- Must-read Papers for File System (FS)☆330Dec 17, 2025Updated 9 months ago
- GeminiFS: A Companion File System for GPUs☆85Aug 11, 2026Updated last month
- InfiniGen: Efficient Generative Inference of Large Language Models with Dynamic KV Cache Management (OSDI'24)☆193Jul 10, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆15Apr 11, 2024Updated 2 years ago
- ☆25Mar 31, 2022Updated 4 years ago
- ☆18May 16, 2022Updated 4 years ago
- NVMe based File System in User-space☆112Feb 16, 2020Updated 6 years ago
- ☆23Jun 21, 2023Updated 3 years ago
- MLPerf® Storage Benchmark Suite☆204Updated this week
- LLM checkpointing for DeepSpeed/Megatron☆27Nov 30, 2025Updated 9 months ago
- ☆23Feb 16, 2023Updated 3 years ago
- Official implementation of ICML 2024 paper "ExCP: Extreme LLM Checkpoint Compression via Weight-Momentum Joint Shrinking".☆48Jul 12, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- CrossFS: A Cross-layered Direct-Access File System☆27Sep 5, 2022Updated 4 years ago
- A mirror of https://bitbucket.org/ajaustin/hemem/src/sosp-submission/☆29May 3, 2023Updated 3 years ago
- This repository is deprecated. See below for details.☆31Jul 27, 2018Updated 8 years ago
- ☆38Jan 15, 2021Updated 5 years ago
- [NeurIPS 2024] Implementation of paper - D-LLM: A Token Adaptive Computing Resource Allocation Strategy for Large Language Models☆26Apr 9, 2025Updated last year
- An I/O benchmark for deep Learning applications☆111Jun 18, 2026Updated 3 months ago
- NEO is a LLM inference engine built to save the GPU memory crisis by CPU offloading☆99Jun 16, 2025Updated last year
- Dynamic Memory Management for Serving LLMs without PagedAttention☆524Aug 24, 2026Updated last month
- A tracing tool to analyze the I/O behavior of a program.☆12Sep 25, 2019Updated 7 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Write a simple file system from zero.☆12Apr 14, 2024Updated 2 years ago
- Source code for the FAST '23 paper “MadFS: Per-File Virtualization for Userspace Persistent Memory Filesystems”☆51Mar 5, 2023Updated 3 years ago
- Profiling and Improving the PyTorch Dataloader for high-latency Storage☆21Apr 18, 2023Updated 3 years ago
- Everything you always wanted to know about ANNS but were afraid to ask 🥰☆12Dec 21, 2024Updated last year
- It's an experiment based on 09 KDD paper, Beyond Blacklists: Learning to Detect Malicious Web Sites from Suspicious URLs☆10Jan 8, 2019Updated 7 years ago
- ☆119May 3, 2021Updated 5 years ago
- Linux Cross-Memory Attach☆24Apr 21, 2026Updated 5 months ago
- Context File System for Agentic AI☆40Sep 9, 2026Updated 2 weeks ago
- RFUSE: Modernizing Userspace Filesystem Framework through Scalable Kernel-Userspace Communication☆76May 8, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A Filesystem Semi-Microkernel.☆48Oct 24, 2023Updated 2 years ago
- ☆26Dec 12, 2017Updated 8 years ago
- KV cache store for distributed LLM inference☆444Sep 16, 2026Updated last week
- Lustre Repository with MS patches☆16Sep 20, 2026Updated last week
- Public repository for the BeeGFS Parallel File System☆229Aug 19, 2026Updated last month
- RecStore: High-performance parameter storage for large-scale recommendation models, unifying heterogeneous memory as a scalable embedding…☆43Updated this week
- ☆204Jul 15, 2025Updated last year