[ICLR 2026π₯] MHLA: Restoring Expressivity of Linear Attention via Token-Level Multi-Head
β157May 19, 2026Updated 3 months ago
Alternatives and similar repositories for MHLA
Users that are interested in MHLA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV2025 highlight]Rectifying Magnitude Neglect in Linear Attentionβ64Jul 24, 2025Updated last year
- [ICML 2026π₯]Rethinking Video Generation Model for the Embodied Worldβ98Jun 1, 2026Updated 3 months ago
- Piecewise Sparse Attention Is Wiser for Efficient Diffusion Transformersβ44Jul 1, 2026Updated 2 months ago
- HumanNet: Scaling Human-centric Video Learning to One Million Hoursβ290May 26, 2026Updated 3 months ago
- Scaling Text-to-Image Diffusion Transformers with Representation Autoencodersβ263Feb 13, 2026Updated 7 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Official implementation of MAGREF: Masked Guidance for Any-Reference Video Generation with Subject Disentanglement (ICLR2026)β299Mar 24, 2026Updated 5 months ago
- [TIP 2026] "FourierSR: A Fourier Token-based Plugin for Efficient Image Super-Resolution"β17Feb 4, 2026Updated 7 months ago
- SLA: Beyond Sparsity in Diffusion Transformers via Fine-Tunable SparseβLinear Attentionβ360Feb 24, 2026Updated 6 months ago
- π· [CVPR'26] Camera-controlled text-to-video generation, now with intrinsics, distortion and orientation control!β231May 15, 2026Updated 3 months ago
- [AAAI 2026] Flowing Backwards: Improving Normalizing Flows via Reverse Representation Alignmentβ17Dec 9, 2025Updated 9 months ago
- [CVPR 2026π₯] Enhancing Spatial Understanding in Image Generation via Reward Modelingβ86Mar 2, 2026Updated 6 months ago
- Minute-long video generation at 24FPS.β70Mar 28, 2026Updated 5 months ago
- [CVPR 2026 Highlight] Official implementation of Log-linear Sparse Attention (LLSA).β92May 1, 2026Updated 4 months ago
- [ICML 2026] Official codebase for "Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactivβ¦β959Aug 28, 2026Updated 2 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Flow Map OPD for AnyStep Video Diffusionβ425Aug 14, 2026Updated 3 weeks ago
- Video models as pure visual reasoners for high-quality text-to-image generation via Chain-of-Frame reasoning.β42Jan 16, 2026Updated 7 months ago
- Official implementation of paper "VMoBA: Mixture-of-Block Attention for Video Diffusion Models"β64Jul 1, 2025Updated last year
- [2025] ModalFormer: Multimodal Transformer for Low-Light Image Enhancementβ28Apr 14, 2026Updated 4 months ago
- Code Implementation of "WorldCam: Interactive Autoregressive 3D Gaming Worlds with Camera Pose as a Unifying Geometric Representation"β183May 9, 2026Updated 4 months ago
- Vision Bridge Transformer at Scaleβ148Dec 1, 2025Updated 9 months ago
- [ICML 2026] "LIVE: Long-horizon Interactive Video World ModEling"β43Jul 15, 2026Updated last month
- [CVPR 2026 Findings] Official repository for "SAT: Selective Aggregation Transformer for Image Super-Resolution"β40Jun 1, 2026Updated 3 months ago
- rCM & Causal-rCM: Leading and Unified Algorithms/Infrastructures for Bidirectional/Autoregressive Video Diffusion Distillation at Scaleβ803Jun 25, 2026Updated 2 months ago
- Managed Kubernetes at scale on DigitalOcean β’ AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [ICML 2026] Stable Velocity: A Variance Perspective on Flow Matchingβ30Feb 19, 2026Updated 6 months ago
- [ICLRβ26] Learning Video Generation for Robotic Manipulation with Collaborative Trajectory Controlβ107Feb 8, 2026Updated 7 months ago
- Official PyTorch implementation of [PSA: Pyramid Sparse Attention for Efficient Video Understanding and Generation](https://arxiv.org/absβ¦β25Jan 25, 2026Updated 7 months ago
- [AAAI 2026] Official repository of Circulant Attentionβ67Jun 26, 2026Updated 2 months ago
- [CVPR2026] VideoAuto-R1: Video Auto Reasoning via Thinking Once, Answering Twiceβ89Feb 27, 2026Updated 6 months ago
- Official Implementation of "MemFlow: Flowing Adaptive Memory for Consistent and Efficient Long Video Narratives"β217Dec 29, 2025Updated 8 months ago
- [NeurIPS 2025] Training-Free Efficient Video Generation via Dynamic Token Carvingβ289Aug 4, 2025Updated last year
- [ICLR 2025] This repo is the official implementation of "The Labyrinth of Links: Navigating the Associative Maze of Multi-modal LLMs".β13Jan 25, 2025Updated last year
- [CVPR 2026] Official repo for "EVATok: Adaptive Length Video Tokenization for Efficient Visual Autoregressive Generation"β65Mar 13, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICML 2026] ReCo: In-Context Generation with Regional Constraints for Instructional Video Editingβ178Aug 20, 2026Updated 3 weeks ago
- PyTorch Implementation of LocAtViT in "Locality-Attending Vision Transformer" (ICLR 2026)β19Aug 12, 2026Updated last month
- StableWorld: Towards Stable and Consistent Long Interactive Video Generationβ101Aug 30, 2026Updated 2 weeks ago
- [ICLR 2026] Official implementation of "Enhancing Multi-Image Understanding Through Delimiter Token Scaling"β17Jul 10, 2026Updated 2 months ago
- Code for "StreamingTalker: Audio-driven 3D Facial Animation with Autoregressive Diffusion Model", AAAI2026 Oralβ60Jun 15, 2026Updated 2 months ago
- This code implements the algorithm of FIPO, a value-free RL recipe for eliciting deeper reasoning from a clean base model.β19Jul 14, 2026Updated last month
- Image-Free Timestep Distillation via Continuous-Time Consistency with Trajectory-Sampled Pairsβ21Dec 16, 2025Updated 8 months ago