Accelerating AI Training and Inference from Storage Perspective (Must-read Papers on Storage for AI)
☆64Aug 16, 2026Updated this week
Alternatives and similar repositories for Storage-for-AI-Paper
Users that are interested in Storage-for-AI-Paper are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The code based on vLLM for the paper “ Cost-Efficient Large Language Model Serving for Multi-turn Conversations with CachedAttention”.☆11Sep 19, 2024Updated last year
- Must-read Papers for File System (FS)☆326Dec 17, 2025Updated 8 months ago
- GeminiFS: A Companion File System for GPUs☆85Aug 11, 2026Updated last week
- ☆15Apr 11, 2024Updated 2 years ago
- ☆25Mar 31, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆18May 16, 2022Updated 4 years ago
- NVMe based File System in User-space☆112Feb 16, 2020Updated 6 years ago
- ☆23Jun 21, 2023Updated 3 years ago
- MLPerf® Storage Benchmark Suite☆196Updated this week
- LLM checkpointing for DeepSpeed/Megatron☆26Nov 30, 2025Updated 8 months ago
- ☆16Apr 13, 2024Updated 2 years ago
- ☆23Feb 16, 2023Updated 3 years ago
- ☆25Aug 11, 2026Updated last week
- A mirror of https://bitbucket.org/ajaustin/hemem/src/sosp-submission/☆28May 3, 2023Updated 3 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- This repository is deprecated. See below for details.☆31Jul 27, 2018Updated 8 years ago
- ☆39Jan 15, 2021Updated 5 years ago
- Decoding Attention is specially optimized for MHA, MQA, GQA and MLA using CUDA core for the decoding stage of LLM inference.☆48Jun 11, 2025Updated last year
- [NeurIPS 2024] Implementation of paper - D-LLM: A Token Adaptive Computing Resource Allocation Strategy for Large Language Models☆25Apr 9, 2025Updated last year
- An I/O benchmark for deep Learning applications☆109Jun 18, 2026Updated 2 months ago
- NEO is a LLM inference engine built to save the GPU memory crisis by CPU offloading☆99Jun 16, 2025Updated last year
- Dynamic Memory Management for Serving LLMs without PagedAttention☆513Jul 17, 2026Updated last month
- A tracing tool to analyze the I/O behavior of a program.☆12Sep 25, 2019Updated 6 years ago
- Write a simple file system from zero.☆12Apr 14, 2024Updated 2 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- A Fast Graph Update Library for FPGA-based Dynamic Graph Processing☆10Dec 20, 2021Updated 4 years ago
- ☆10Feb 22, 2023Updated 3 years ago
- Gengar, a distributed shared hybrid memory pool with RDMA support. Gengar allows applications to access remote DRAM/NVM in a large and gl…☆24May 24, 2022Updated 4 years ago
- Source code for the FAST '23 paper “MadFS: Per-File Virtualization for Userspace Persistent Memory Filesystems”☆51Mar 5, 2023Updated 3 years ago
- Profiling and Improving the PyTorch Dataloader for high-latency Storage☆21Apr 18, 2023Updated 3 years ago
- It's an experiment based on 09 KDD paper, Beyond Blacklists: Learning to Detect Malicious Web Sites from Suspicious URLs☆10Jan 8, 2019Updated 7 years ago
- ☆119May 3, 2021Updated 5 years ago
- Context File System for Agentic AI☆39Jul 14, 2026Updated last month
- RFUSE: Modernizing Userspace Filesystem Framework through Scalable Kernel-Userspace Communication☆75May 8, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆26Dec 12, 2017Updated 8 years ago
- KV cache store for distributed LLM inference☆427Nov 13, 2025Updated 9 months ago
- Public repository for the BeeGFS Parallel File System☆226Jul 15, 2026Updated last month
- RecStore: High-performance parameter storage for large-scale recommendation models, unifying heterogeneous memory as a scalable embedding…☆43Updated this week
- ☆202Jul 15, 2025Updated last year
- Human-initiated, AI-powered knowledge graph for markdown documentation☆15Jan 30, 2026Updated 6 months ago
- This repository serves as a comprehensive survey of LLM development, featuring numerous research papers along with their corresponding co…☆350Jul 16, 2026Updated last month