Implementation for paper "Forcing-KV: Hybrid KV Cache Compression for Efficient Autoregressive Video Diffusion Models".
☆118May 17, 2026Updated 2 months ago
Alternatives and similar repositories for Forcing-KV
Users that are interested in Forcing-KV are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official page of ImmerIris: A Large-Scale Dataset and Benchmark for Off-Axis and Unconstrained Iris Recognition in Immersive Applications…☆30Jun 9, 2026Updated last month
- [ACL 2026 Main] See the Forest for the Trees: Loosely Speculative Decoding via Visual-Semantic Guidance for Efficient Inference of Video …☆27Jul 4, 2026Updated 2 weeks ago
- Training Code for ADMIS Teams in CVPR2024 FRCSyn Competition☆33Jan 5, 2026Updated 6 months ago
- Implementation of our paper "RAVEN: Real-time Autoregressive Video Extrapolation with Consistency-model GRPO"☆52Jul 11, 2026Updated last week
- [EMNLP 2025 Main] SpecVLM: Enhancing Speculative Decoding of Video LLMs via Verifier-Guided Token Pruning☆48Apr 16, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- AnyTalker: Scaling Multi-person Talking Video Generation with Interactivity Refinement☆322Apr 15, 2026Updated 3 months ago
- Official repository for: IDDR-NGP:Incorporating Detectors for Distractors Removal with Instant Neural Radiance Field☆23Jul 8, 2024Updated 2 years ago
- Minute-long video generation at 24FPS.☆69Mar 28, 2026Updated 3 months ago
- The official code of "Beyond VLM-Based Rewards: Diffusion-Native Latent Reward Modeling"☆66Jun 30, 2026Updated 3 weeks ago
- EchoStyle: Unlocking High-Fidelity Video Stylization with Reverse Data Synthesis☆29Jul 2, 2026Updated 2 weeks ago
- [ACL 2026 Findings] Living repository for the survey paper “Efficient Inference for Large Vision-Language Models: Bottlenecks, Techniques…☆26Apr 8, 2026Updated 3 months ago
- [CVPR2026] VecAttention: Vector-wise Sparse Attention for Accelerating Long-Context Inference☆20May 27, 2026Updated last month
- ☆19Feb 18, 2025Updated last year
- Official implementation of "EponaV2: Driving World Model with Comprehensive Future Reasoning"☆33May 14, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [CVPR 2026] Action-Controllable Infinite Video Generation Emerges From Autoregressive Self-Rollout☆86Mar 21, 2026Updated 4 months ago
- [ICML 2026] Official implementation of "Deep Forcing: Training-Free Long Video Generation with Deep Sink and Participative Compression"☆135Apr 30, 2026Updated 2 months ago
- ☆15Jan 27, 2026Updated 5 months ago
- [ICML2026] Auto-Regressive Long Video Generation via 2-Bit KV-Cache Quantization☆60Jun 4, 2026Updated last month
- [ECCV 2026] Official implementation of "MemRoPE: Training-Free Infinite Video Generation via Evolving Memory Tokens"☆46Jun 24, 2026Updated 3 weeks ago
- ☆37Feb 12, 2026Updated 5 months ago
- [CVPR 2026 Highlight] Official implementation of BiCo: Composing Concepts from Images and Videos via Concept-prompt Binding☆85May 31, 2026Updated last month
- [ICML 2026] Official codebase for "Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactiv…☆869Jul 9, 2026Updated last week
- Official implementation of "Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion"☆104Jul 8, 2026Updated last week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A list of awesome papers on compression and acceleration of Large Language Models (LLMs) and Multimodal Large Language Models (MLLMs).☆16May 12, 2026Updated 2 months ago
- A Minimal and Elegant Framework & Tutorial for Real-Time Interactive World Models☆723Jun 15, 2026Updated last month
- [AAAI 2024] MLNet: Mutual Learning Network with Neighborhood Invariance for Universal Domain Adaptation☆21Feb 29, 2024Updated 2 years ago
- [CVPR 2024 Highlight] Coarse-to-Fine Latent Diffusion for Pose-Guided Person Image Synthesis☆243Jan 22, 2025Updated last year
- Official Repo for the VideoVerse☆15Mar 29, 2026Updated 3 months ago
- A curated list of recent papers on efficient video attention for video diffusion models, including sparsification, quantization, and cach…☆61Oct 27, 2025Updated 8 months ago
- DomainShuttle: Freeform Open Domain Subject-driven Text-to-video Generation☆162Jun 26, 2026Updated 3 weeks ago
- ☆54May 6, 2026Updated 2 months ago
- ☆20May 21, 2026Updated 2 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ICLR 2026] Uni-CoT: Towards Unified Chain-of-Thought Reasoning Across Text and Vision☆233May 31, 2026Updated last month
- A Curated List of Awesome Video World Models with AR Diffusion: Covering Algorithms, Applications, and Infrastructure, Aimed at Serving a…☆659Jun 4, 2026Updated last month
- [ICLR 2026] Official Repo for Rolling Forcing: Autoregressive Long Video Diffusion in Real Time☆442Oct 31, 2025Updated 8 months ago
- The official implementation of MotionGrasp☆38Nov 15, 2025Updated 8 months ago
- [ICML 2026] Official repository for the paper "Light Forcing: Accelerating Autoregressive Video Diffusion via Sparse Attention"☆41May 24, 2026Updated last month
- Infinite-Forcing: Towards Infinite-Long Video Generation☆155Nov 13, 2025Updated 8 months ago
- Code repository for the SOSP'25 paper DCP: Addressing Input Dynamism In Long-Context Training via Dynamic Context Parallelism.☆21Nov 28, 2025Updated 7 months ago