Accelerating AI Training and Inference from Storage Perspective (Must-read Papers on Storage for AI)
☆64Aug 16, 2026Updated 3 weeks ago
Alternatives and similar repositories for Storage-for-AI-Paper
Users that are interested in Storage-for-AI-Paper are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The code based on vLLM for the paper “ Cost-Efficient Large Language Model Serving for Multi-turn Conversations with CachedAttention”.☆11Sep 19, 2024Updated last year
- ☆15Jan 21, 2023Updated 3 years ago
- Must-read Papers for File System (FS)☆327Dec 17, 2025Updated 8 months ago
- GeminiFS: A Companion File System for GPUs☆86Aug 11, 2026Updated 3 weeks ago
- InfiniGen: Efficient Generative Inference of Large Language Models with Dynamic KV Cache Management (OSDI'24)☆192Jul 10, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆15Apr 11, 2024Updated 2 years ago
- ☆25Mar 31, 2022Updated 4 years ago
- NVMe based File System in User-space☆112Feb 16, 2020Updated 6 years ago
- ☆23Jun 21, 2023Updated 3 years ago
- DL Dataloader Benchmarks☆20Jan 27, 2025Updated last year
- LLM checkpointing for DeepSpeed/Megatron☆26Nov 30, 2025Updated 9 months ago
- ☆16Apr 13, 2024Updated 2 years ago
- ☆23Feb 16, 2023Updated 3 years ago
- CrossFS: A Cross-layered Direct-Access File System☆27Sep 5, 2022Updated 4 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A mirror of https://bitbucket.org/ajaustin/hemem/src/sosp-submission/☆28May 3, 2023Updated 3 years ago
- This repository is deprecated. See below for details.☆31Jul 27, 2018Updated 8 years ago
- Decoding Attention is specially optimized for MHA, MQA, GQA and MLA using CUDA core for the decoding stage of LLM inference.☆48Jun 11, 2025Updated last year
- [NeurIPS 2024] Implementation of paper - D-LLM: A Token Adaptive Computing Resource Allocation Strategy for Large Language Models☆25Apr 9, 2025Updated last year
- An I/O benchmark for deep Learning applications☆110Jun 18, 2026Updated 2 months ago
- NEO is a LLM inference engine built to save the GPU memory crisis by CPU offloading☆99Jun 16, 2025Updated last year
- Write a simple file system from zero.☆12Apr 14, 2024Updated 2 years ago
- A Fast Graph Update Library for FPGA-based Dynamic Graph Processing☆10Dec 20, 2021Updated 4 years ago
- ☆10Feb 22, 2023Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Gengar, a distributed shared hybrid memory pool with RDMA support. Gengar allows applications to access remote DRAM/NVM in a large and gl…☆24May 24, 2022Updated 4 years ago
- Source code for the FAST '23 paper “MadFS: Per-File Virtualization for Userspace Persistent Memory Filesystems”☆51Mar 5, 2023Updated 3 years ago
- Profiling and Improving the PyTorch Dataloader for high-latency Storage☆21Apr 18, 2023Updated 3 years ago
- Everything you always wanted to know about ANNS but were afraid to ask 🥰☆12Dec 21, 2024Updated last year
- It's an experiment based on 09 KDD paper, Beyond Blacklists: Learning to Detect Malicious Web Sites from Suspicious URLs☆10Jan 8, 2019Updated 7 years ago
- ☆119May 3, 2021Updated 5 years ago
- Context File System for Agentic AI☆39Jul 14, 2026Updated last month
- KV cache store for distributed LLM inference☆436Nov 13, 2025Updated 9 months ago
- Public repository for the BeeGFS Parallel File System☆226Aug 19, 2026Updated 2 weeks ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Lustre Repository with MS patches☆17Updated this week
- RecStore: High-performance parameter storage for large-scale recommendation models, unifying heterogeneous memory as a scalable embedding…☆43Updated this week
- ☆203Jul 15, 2025Updated last year
- Human-initiated, AI-powered knowledge graph for markdown documentation☆15Jan 30, 2026Updated 7 months ago
- This repository serves as a comprehensive survey of LLM development, featuring numerous research papers along with their corresponding co…☆355Jul 16, 2026Updated last month
- Experimental repository for GSoC 2024.☆15Aug 29, 2024Updated 2 years ago
- ☆44Jul 10, 2017Updated 9 years ago