Implementation for paper "Forcing-KV: Hybrid KV Cache Compression for Efficient Autoregressive Video Diffusion Models".
☆138May 17, 2026Updated 4 months ago
Alternatives and similar repositories for Forcing-KV
Users that are interested in Forcing-KV are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official page of ImmerIris: A Large-Scale Dataset and Benchmark for Off-Axis and Unconstrained Iris Recognition in Immersive Applications…☆29Jun 9, 2026Updated 4 months ago
- [ACL 2026 Main] See the Forest for the Trees: Loosely Speculative Decoding via Visual-Semantic Guidance for Efficient Inference of Video …☆30Jul 4, 2026Updated 3 months ago
- [EMNLP 2025 Main] SpecVLM: Enhancing Speculative Decoding of Video LLMs via Verifier-Guided Token Pruning☆49Apr 16, 2026Updated 5 months ago
- [NeurIPS 2026] RAVEN: Real-time Autoregressive Video Extrapolation with Consistency-model GRPO☆134Sep 29, 2026Updated last week
- Official repository for: IDDR-NGP:Incorporating Detectors for Distractors Removal with Instant Neural Radiance Field☆23Jul 8, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- AnyTalker: Scaling Multi-person Talking Video Generation with Interactivity Refinement☆326Apr 15, 2026Updated 5 months ago
- ☆19Feb 18, 2025Updated last year
- Minute-long video generation at 24FPS.☆70Mar 28, 2026Updated 6 months ago
- The official code of "Beyond VLM-Based Rewards: Diffusion-Native Latent Reward Modeling"☆68Jun 30, 2026Updated 3 months ago
- An out-of-the-box inference acceleration engine for Diffusion and DiT models☆58Mar 21, 2025Updated last year
- EchoStyle: Unlocking High-Fidelity Video Stylization with Reverse Data Synthesis☆35Sep 7, 2026Updated last month
- [ACL 2026 Findings] Living repository for the survey paper “Efficient Inference for Large Vision-Language Models: Bottlenecks, Techniques…☆28Sep 22, 2026Updated 2 weeks ago
- [CVPR2026] VecAttention: Vector-wise Sparse Attention for Accelerating Long-Context Inference☆22May 27, 2026Updated 4 months ago
- Official implementation of "EponaV2: Driving World Model with Comprehensive Future Reasoning"☆35May 14, 2026Updated 4 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- [ICML 2026] Official implementation of "Deep Forcing: Training-Free Long Video Generation with Deep Sink and Participative Compression"☆148Apr 30, 2026Updated 5 months ago
- [CVPR 2026] Action-Controllable Infinite Video Generation Emerges From Autoregressive Self-Rollout☆100Mar 21, 2026Updated 6 months ago
- ☆14Jan 27, 2026Updated 8 months ago
- [ICML2026] Auto-Regressive Long Video Generation via 2-Bit KV-Cache Quantization☆67Jul 26, 2026Updated 2 months ago
- [ECCV 2026] Official implementation of "MemRoPE: Training-Free Infinite Video Generation via Evolving Memory Tokens"☆60Jun 24, 2026Updated 3 months ago
- ☆38Feb 12, 2026Updated 7 months ago
- [CVPR 2026 Highlight] Official implementation of BiCo: Composing Concepts from Images and Videos via Concept-prompt Binding☆87May 31, 2026Updated 4 months ago
- A Minimal and Elegant Framework & Tutorial for Real-Time Interactive World Models☆851Sep 10, 2026Updated last month
- Official implementation of "Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion"☆108Aug 25, 2026Updated last month
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [ICML 2026] Official codebase for "Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactiv…☆999Oct 3, 2026Updated last week
- A list of awesome papers on compression and acceleration of Large Language Models (LLMs) and Multimodal Large Language Models (MLLMs).☆18May 12, 2026Updated 4 months ago
- [AAAI 2024] MLNet: Mutual Learning Network with Neighborhood Invariance for Universal Domain Adaptation☆21Feb 29, 2024Updated 2 years ago
- [CVPR 2024 Highlight] Coarse-to-Fine Latent Diffusion for Pose-Guided Person Image Synthesis☆242Jan 22, 2025Updated last year
- Official Repo for the VideoVerse☆15Mar 29, 2026Updated 6 months ago
- A curated list of recent papers on efficient video attention for video diffusion models, including sparsification, quantization, and cach…☆65Oct 27, 2025Updated 11 months ago
- DomainShuttle: Freeform Open Domain Subject-driven Text-to-video Generation☆168Jun 26, 2026Updated 3 months ago
- ☆56May 6, 2026Updated 5 months ago
- ☆20May 21, 2026Updated 4 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [ICLR 2026] Uni-CoT: Towards Unified Chain-of-Thought Reasoning Across Text and Vision☆237May 31, 2026Updated 4 months ago
- [ICLR 2026] Official Repo for Rolling Forcing: Autoregressive Long Video Diffusion in Real Time☆461Oct 31, 2025Updated 11 months ago
- A Curated List of Awesome Video World Models with AR Diffusion: Covering Algorithms, Applications, and Infrastructure, Aimed at Serving a…☆715Updated this week
- [ICML 2026] Official repository for the paper "Light Forcing: Accelerating Autoregressive Video Diffusion via Sparse Attention"☆53Aug 5, 2026Updated 2 months ago
- The official implementation of MotionGrasp☆40Nov 15, 2025Updated 10 months ago
- A portable and efficient infrastracture for value profilers. Doc: https://vclinic.readthedocs.io/en/latest/index.html☆14Mar 4, 2026Updated 7 months ago
- Code repository for the SOSP'25 paper DCP: Addressing Input Dynamism In Long-Context Training via Dynamic Context Parallelism.☆23Nov 28, 2025Updated 10 months ago