[ICLR 2026π₯] MHLA: Restoring Expressivity of Linear Attention via Token-Level Multi-Head
β154May 19, 2026Updated 3 months ago
Alternatives and similar repositories for MHLA
Users that are interested in MHLA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV2025 highlight]Rectifying Magnitude Neglect in Linear Attentionβ64Jul 24, 2025Updated last year
- [ICML 2026π₯]Rethinking Video Generation Model for the Embodied Worldβ93Jun 1, 2026Updated 2 months ago
- Piecewise Sparse Attention Is Wiser for Efficient Diffusion Transformersβ44Jul 1, 2026Updated last month
- HumanNet: Scaling Human-centric Video Learning to One Million Hoursβ283May 26, 2026Updated 2 months ago
- Scaling Text-to-Image Diffusion Transformers with Representation Autoencodersβ258Feb 13, 2026Updated 6 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Official implementation of MAGREF: Masked Guidance for Any-Reference Video Generation with Subject Disentanglement (ICLR2026)β300Mar 24, 2026Updated 5 months ago
- [TIP 2026] "FourierSR: A Fourier Token-based Plugin for Efficient Image Super-Resolution"β17Feb 4, 2026Updated 6 months ago
- SLA: Beyond Sparsity in Diffusion Transformers via Fine-Tunable SparseβLinear Attentionβ339Feb 24, 2026Updated 6 months ago
- π· [CVPR'26] Camera-controlled text-to-video generation, now with intrinsics, distortion and orientation control!β222May 15, 2026Updated 3 months ago
- [AAAI 2026] Flowing Backwards: Improving Normalizing Flows via Reverse Representation Alignmentβ16Dec 9, 2025Updated 8 months ago
- [CVPR 2026π₯] Enhancing Spatial Understanding in Image Generation via Reward Modelingβ86Mar 2, 2026Updated 5 months ago
- Minute-long video generation at 24FPS.β69Mar 28, 2026Updated 4 months ago
- [CVPR 2026 Highlight] Official implementation of Log-linear Sparse Attention (LLSA).β93May 1, 2026Updated 3 months ago
- [ICML 2026] Official codebase for "Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactivβ¦β932Jul 23, 2026Updated last month
- Deploy open-source AI quickly and easily - Special Bonus Offer β’ AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Flow Map OPD for AnyStep Video Diffusionβ413Aug 14, 2026Updated last week
- Video models as pure visual reasoners for high-quality text-to-image generation via Chain-of-Frame reasoning.β40Jan 16, 2026Updated 7 months ago
- Official implementation of paper "VMoBA: Mixture-of-Block Attention for Video Diffusion Models"β64Jul 1, 2025Updated last year
- [2025] ModalFormer: Multimodal Transformer for Low-Light Image Enhancementβ28Apr 14, 2026Updated 4 months ago
- Code Implementation of "WorldCam: Interactive Autoregressive 3D Gaming Worlds with Camera Pose as a Unifying Geometric Representation"β180May 9, 2026Updated 3 months ago
- Vision Bridge Transformer at Scaleβ148Dec 1, 2025Updated 8 months ago
- [ICML 2026] "LIVE: Long-horizon Interactive Video World ModEling"β42Jul 15, 2026Updated last month
- [CVPR 2026 Findings] Official repository for "SAT: Selective Aggregation Transformer for Image Super-Resolution"β40Jun 1, 2026Updated 2 months ago
- rCM & Causal-rCM: Leading and Unified Algorithms/Infrastructures for Bidirectional/Autoregressive Video Diffusion Distillation at Scaleβ788Jun 25, 2026Updated last month
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [ICML 2026] Stable Velocity: A Variance Perspective on Flow Matchingβ30Feb 19, 2026Updated 6 months ago
- [ICLRβ26] Learning Video Generation for Robotic Manipulation with Collaborative Trajectory Controlβ107Feb 8, 2026Updated 6 months ago
- Official PyTorch implementation of [PSA: Pyramid Sparse Attention for Efficient Video Understanding and Generation](https://arxiv.org/absβ¦β25Jan 25, 2026Updated 6 months ago
- [AAAI 2026] Official repository of Circulant Attentionβ65Jun 26, 2026Updated last month
- [CVPR2026] VideoAuto-R1: Video Auto Reasoning via Thinking Once, Answering Twiceβ88Feb 27, 2026Updated 5 months ago
- Official Implementation of "MemFlow: Flowing Adaptive Memory for Consistent and Efficient Long Video Narratives"β216Dec 29, 2025Updated 7 months ago
- [NeurIPS 2025] Training-Free Efficient Video Generation via Dynamic Token Carvingβ289Aug 4, 2025Updated last year
- [ICLR 2025] This repo is the official implementation of "The Labyrinth of Links: Navigating the Associative Maze of Multi-modal LLMs".β13Jan 25, 2025Updated last year
- [CVPR 2026] Official repo for "EVATok: Adaptive Length Video Tokenization for Efficient Visual Autoregressive Generation"β65Mar 13, 2026Updated 5 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits β’ AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ICML 2026] ReCo: In-Context Generation with Regional Constraints for Instructional Video Editingβ176Updated this week
- PyTorch Implementation of LocAtViT in "Locality-Attending Vision Transformer" (ICLR 2026)β19Aug 12, 2026Updated last week
- StableWorld: Towards Stable and Consistent Long Interactive Video Generationβ97Mar 18, 2026Updated 5 months ago
- [ICLR 2026] Official implementation of "Enhancing Multi-Image Understanding Through Delimiter Token Scaling"β17Jul 10, 2026Updated last month
- Code for "StreamingTalker: Audio-driven 3D Facial Animation with Autoregressive Diffusion Model", AAAI2026 Oralβ58Jun 15, 2026Updated 2 months ago
- This code implements the algorithm of FIPO, a value-free RL recipe for eliciting deeper reasoning from a clean base model.β18Jul 14, 2026Updated last month
- Image-Free Timestep Distillation via Continuous-Time Consistency with Trajectory-Sampled Pairsβ21Dec 16, 2025Updated 8 months ago