Aiming to integrate most existing feature caching-based diffusion acceleration schemes into a unified framework.
☆111Oct 23, 2025Updated 9 months ago
Alternatives and similar repositories for Cache4Diffusion
Users that are interested in Cache4Diffusion are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A curated list of research papers, resources, and advancements on Diffusion Cache and related efficient diffusion model acceleration tech…☆89Jul 23, 2026Updated 2 weeks ago
- [ICCV2025] From Reusing to Forecasting: Accelerating Diffusion Models with TaylorSeers☆410Mar 2, 2026Updated 5 months ago
- ☆25Sep 4, 2025Updated 11 months ago
- [ICLR2025] Accelerating Diffusion Transformers with Token-wise Feature Caching☆221Mar 14, 2025Updated last year
- [ICLR 2026] Official implementation of DiCache: Let Diffusion Model Determine Its Own Cache☆62Jan 26, 2026Updated 6 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [CVPR 2026 Oral, Best Paper Finalist] SeaCache: Spectral-Evolution-Aware Cache for Accelerating Diffusion Models☆96Jun 29, 2026Updated last month
- [CVPR 2026] Denoising as Path Planning: Training-Free Acceleration of Diffusion Models with DPCache☆43Jul 1, 2026Updated last month
- HiCache: Hermite Polynomial-based Feature Cache for diffusion inference☆15Jul 29, 2026Updated last week
- [CVPR 2026 Oral] SenCache: Accelerating Diffusion Model Inference via Sensitivity-Aware Caching☆24Jun 5, 2026Updated 2 months ago
- 📚 Collection of awesome generation acceleration resources.☆402Jul 7, 2025Updated last year
- DisCa: Accelerating Video Diffusion Transformers with Distillation-Compatible Learnable Feature Caching☆24Apr 15, 2026Updated 3 months ago
- ☆22Nov 3, 2025Updated 9 months ago
- A PyTorch-native inference engine with cache, parallelism, quantization and cpu offload for DiTs.☆1,243Jul 26, 2026Updated 2 weeks ago
- FastCache: Fast Caching for Diffusion Transformer Through Learnable Linear Approximation [Efficient ML Model]☆52Aug 2, 2026Updated last week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ICLR2026] The open-source code for FlowCache, including accelerated implementations of the MAGI-1 and Skyreels-V2.☆30Apr 24, 2026Updated 3 months ago
- 📚A curated list of Awesome Diffusion Inference Papers with Codes: Sampling, Cache, Quantization, Parallelism, etc.🎉☆581Jun 13, 2026Updated last month
- [ICML2026] Auto-Regressive Long Video Generation via 2-Bit KV-Cache Quantization☆61Jul 26, 2026Updated 2 weeks ago
- [NeurIPS 2025 Spotlight] LeMiCa: Lexicographic Minimax Path Caching for Efficient Diffusion-Based Video Generation☆122Jun 22, 2026Updated last month
- 🎬 3.7× faster video generation E2E 🖼️ 1.6× faster image generation E2E ⚡ ColumnSparseAttn 9.3× vs FlashAttn‑3 💨 ColumnSparseGEMM 2.5× …☆111Sep 8, 2025Updated 11 months ago
- Timestep Embedding Tells: It's Time to Cache for Video Diffusion Model☆1,365Jun 8, 2025Updated last year
- [ICML2025, NeurIPS2025 Spotlight] Sparse VideoGen 1 & 2: Accelerating Video Diffusion Transformers with Sparse Attention☆700Jul 4, 2026Updated last month
- Official ConvRot implementation. A plug-and-play, convolution-like rotation module enabling efficient W4A4 quantization for diffusion mod…☆23Jul 3, 2026Updated last month
- [AAAI-2025] The offical code for SiTo (Similarity-based Token Pruning for Stable Diffusion Models)☆46Jun 2, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [ICLR'25] ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation☆167Mar 21, 2025Updated last year
- [NeurIPS 2025] Training-Free Efficient Video Generation via Dynamic Token Carving☆289Aug 4, 2025Updated last year
- (ICCV2025) EEdit⚡: Rethinking the Spatial and Temporal Redundancy for Efficient Image Editing☆63Sep 17, 2025Updated 10 months ago
- ☆192Jan 14, 2025Updated last year
- Official PyTorch implementation of the paper "dLLM-Cache: Accelerating Diffusion Large Language Models with Adaptive Caching" (dLLM-Cache…☆213May 1, 2026Updated 3 months ago
- [ICML 2026] Official implementation of "NaviCache: Test-Time Self-Calibration Caching for Video Generation".☆27Updated this week
- FORA introduces simple yet effective caching mechanism in Diffusion Transformer Architecture for faster inference sampling.☆56Jul 8, 2024Updated 2 years ago
- Collection of Acceleration Methods for Generative AI☆29Dec 9, 2025Updated 8 months ago
- (NeurIPS 2025 🔥) Official implementation for "Efficient Multi-modal Large Language Models via Progressive Consistency Distillation"☆49Feb 11, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICML2025] SpargeAttention: A training-free sparse attention that accelerates any model inference.☆1,024Feb 25, 2026Updated 5 months ago
- [ICCV 2025] CHORDS: Diffusion Sampling Accelerator with Multi-core Hierarchical ODE Solvers☆17Mar 3, 2026Updated 5 months ago
- Fully Open-source Multimodal Language Models for Science Discovery☆167Mar 20, 2026Updated 4 months ago
- [ICLR 2026] This is the official PyTorch implementation of "QVGen: Pushing the Limit of Quantized Video Generative Models".☆32Feb 11, 2026Updated 5 months ago
- [SIGGRAPH 2026] EasyVFX: This repo is the official implementation of "EasyVFX: Frequency-Driven Decoupling for Resource-Efficient VFX Gen…☆22May 24, 2026Updated 2 months ago
- Code for paper: "Learning Diffusion Models with Flexible Representation Guidance"☆16Mar 18, 2026Updated 4 months ago
- rCM & Causal-rCM: Leading and Unified Algorithms/Infrastructures for Bidirectional/Autoregressive Video Diffusion Distillation at Scale☆780Jun 25, 2026Updated last month