Combining Teacache with xDiT to Accelerate Visual Generation Models
☆33Apr 21, 2025Updated last year
Alternatives and similar repositories for Teacache-xDiT
Users that are interested in Teacache-xDiT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- https://wavespeed.ai/ Context parallel attention that accelerates DiT model inference with dynamic caching☆430Jul 5, 2025Updated last year
- This project is based on the [LTX-Video](https://github.com/Lightricks/LTX-Video) algorithm of the diffusers and optimized and accelerate…☆15Dec 31, 2024Updated last year
- [CVPR 2026] Denoising as Path Planning: Training-Free Acceleration of Diffusion Models with DPCache☆43Jul 1, 2026Updated last month
- [ICML2025, NeurIPS2025 Spotlight] Sparse VideoGen 1 & 2: Accelerating Video Diffusion Transformers with Sparse Attention☆702Jul 4, 2026Updated last month
- A forked version of flux-fast that makes flux-fast even faster with cache-dit, 3.3x speedup on NVIDIA L20.☆24Jul 18, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- An out-of-the-box inference acceleration engine for Diffusion and DiT models☆59Mar 21, 2025Updated last year
- Timestep Embedding Tells: It's Time to Cache for Video Diffusion Model☆1,368Jun 8, 2025Updated last year
- A parallelism VAE avoids OOM for high resolution image generation☆95Updated this week
- Fast and memory-efficient exact attention☆23Jun 26, 2026Updated last month
- ☆25Sep 4, 2025Updated 11 months ago
- ☆431Aug 12, 2026Updated last week
- KsanaDiT: High-Performance DiT (Diffusion Transformer) Inference Framework for Video & Image Generation☆62May 13, 2026Updated 3 months ago
- The official implementation of PTQD: Accurate Post-Training Quantization for Diffusion Models☆103Mar 12, 2024Updated 2 years ago
- [CVPR 2024] DeepCache: Accelerating Diffusion Models for Free☆970Jun 27, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [NeurIPS 2025] Radial Attention: O(nlogn) Sparse Attention with Energy Decay for Long Video Generation☆608Nov 11, 2025Updated 9 months ago
- 📚A curated list of Awesome Diffusion Inference Papers with Codes: Sampling, Cache, Quantization, Parallelism, etc.🎉☆585Jun 13, 2026Updated 2 months ago
- rishipython / One-Eye-is-All-You-Need-Lightweight-Ensembles-for-Gaze-Estimation-with-Single-Encoders☆11Sep 30, 2022Updated 3 years ago
- 2018-2024 in-depth completion of top papers, open source code summary! (Continuous update)☆14Sep 1, 2024Updated last year
- Code for Draft Attention☆103May 22, 2025Updated last year
- Communication-Efficient Diffusion Denoising Parallelization via Reuse-then-Predict Mechanism (NIPS'25)☆16Oct 6, 2025Updated 10 months ago
- [CVPR 2025] T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation☆123Oct 25, 2025Updated 9 months ago
- To pioneer training long-context multi-modal transformer models☆76Aug 8, 2025Updated last year
- ☆10Jun 27, 2018Updated 8 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Unofficial extension implementation of Self-Forcing to support I2V && 14B training.☆385Sep 29, 2025Updated 10 months ago
- [ICCV2025] From Reusing to Forecasting: Accelerating Diffusion Models with TaylorSeers☆411Mar 2, 2026Updated 5 months ago
- [NeurIPS 2024] AsyncDiff: Parallelizing Diffusion Models by Asynchronous Denoising☆214Sep 27, 2025Updated 10 months ago
- Controlnet module for Wan2.2☆45Oct 30, 2025Updated 9 months ago
- A unified inference and post-training framework for accelerated video generation.☆4,029Updated this week
- [ICLR 2025] FasterCache: Training-Free Video Diffusion Model Acceleration with High Quality☆267Dec 27, 2024Updated last year
- FastCache: Fast Caching for Diffusion Transformer Through Learnable Linear Approximation [Efficient ML Model]☆52Aug 2, 2026Updated 3 weeks ago
- ArcFlow: Unleashing 2-Step Text-to-Image Generation via High-Precision Non-Linear Flow Distillation☆130May 20, 2026Updated 3 months ago
- UnitBox: An Advanced Object Detection Network☆23Feb 8, 2017Updated 9 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [CVPR 2023] Domain Generalized Stereo Matching via Hierarchical Visual Transformation☆15Dec 17, 2023Updated 2 years ago
- Cluster Far Mem, framework to execute single job and multi job experiments using fastswap☆21Jan 12, 2024Updated 2 years ago
- ☆66Oct 25, 2025Updated 9 months ago
- Triton kernels for Flux☆23Jul 7, 2025Updated last year
- [ICLR'25] ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation☆167Mar 21, 2025Updated last year
- [ICML2025] SpargeAttention: A training-free sparse attention that accelerates any model inference.☆1,032Feb 25, 2026Updated 5 months ago
- A PyTorch-native inference engine with cache, parallelism, quantization and cpu offload for DiTs.☆1,255Updated this week