[CVPR'26] QuantVLA: Scale-Calibrated Post-Training Quantization for Vision-Language-Action Models
☆47Mar 25, 2026Updated 6 months ago
Alternatives and similar repositories for QuantVLA
Users that are interested in QuantVLA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [NAACL 2025🔥] MEDA: Dynamic KV Cache Allocation for Efficient Multimodal Long-Context Inference☆22Jun 19, 2025Updated last year
- The official implementation of VLA-Pruner: Temporal-Aware Dual-Level Visual Token Pruning for Efficient Vision-Language-Action Inference.☆84Jun 4, 2026Updated 3 months ago
- [ICLR 2026] Mixing Importance with Diversity: Joint Optimization for KV Cache Compression in Large Vision-Language Models☆31Mar 21, 2026Updated 6 months ago
- A performance analysis tool for VLA models☆93Feb 26, 2026Updated 7 months ago
- [ICLR25] STBLLM: Breaking the 1-Bit Barrier with Structured Binary LLMs☆19Jun 3, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [ACM MM'26 Oral] MMaDA-VLA: Large Diffusion Vision-Language-Action Model with Unified Multi-Modal Instruction and Generation☆68May 14, 2026Updated 4 months ago
- [NeurIPS 2023] Token-Scaled Logit Distillation for Ternary Weight Generative Language Models☆18Dec 6, 2023Updated 2 years ago
- [NeurIPS 2026] FASTER: Rethinking Real-Time Flow VLAs☆157May 14, 2026Updated 4 months ago
- [ECCV 2026] Official implementation of CEED-VLA: Consistency Vision-Language-Action Model with Early-Exit Decoding.☆53Sep 15, 2025Updated last year
- Official Code for LightVLA (ICRA 2026)☆105Jan 31, 2026Updated 8 months ago
- [ICLR 2026] Official implementation of "Enhancing Multi-Image Understanding Through Delimiter Token Scaling"☆17Jul 10, 2026Updated 2 months ago
- [CVPR 2026] MergeVLA: Cross-Skill Model Merging Toward a Generalist Vision-Language-Action Agent☆39Aug 19, 2026Updated last month
- [ICLR 2026 Oral] RAIN-Merging☆16Mar 9, 2026Updated 6 months ago
- The code repository of "MBQ: Modality-Balanced Quantization for Large Vision-Language Models"☆97Mar 17, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆23Mar 31, 2026Updated 6 months ago
- [CVPR 2026 Highlight] ForeAct: Steering Your VLA with Efficient Visual Foresight Planning☆96May 1, 2026Updated 5 months ago
- This repository contains the **official implementation** of the paper: "VL2Lite: Task-Specific Knowledge Distillation from Large Vision-…☆21Mar 23, 2025Updated last year
- Official implementation of EMNLP'23 paper "Revisiting Block-based Quantisation: What is Important for Sub-8-bit LLM Inference?"☆24Oct 25, 2023Updated 2 years ago
- [IJCAI 2025] Offical implementation of the paper "Multi-View Learning with Context-Guided Receptance for Image Denoising".☆13Jun 26, 2025Updated last year
- MMDeepResearch-Bench (MMDR)☆33Aug 10, 2026Updated last month
- ☆15Apr 6, 2026Updated 5 months ago
- Official implementation for BitVLA: 1-bit Vision-Language-Action Models for Robotics Manipulation☆167Sep 6, 2026Updated 3 weeks ago
- [ICCV 2025] SparseMM: Head Sparsity Emerges from Visual Concept Responses in MLLMs☆89Jan 17, 2026Updated 8 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [CVPR 2026] Official Implementation for Global Prior Meets Local Consistency: Dual-Memory Augmented Vision-Language-Action Model for Effi…☆34Jul 16, 2026Updated 2 months ago
- 🔥 A curated roadmap to the Efficient VLA landscape. We’re keeping this list live—contribute your latest work!☆213Aug 17, 2026Updated last month
- The Source Code for OmniVideoBench @ICLR 2026☆78Feb 12, 2026Updated 7 months ago
- [NeurIPS 2025] VLA-Cache: Efficient Vision-Language-Action Manipulation via Adaptive Token Caching☆103Feb 27, 2026Updated 7 months ago
- [ICLR 2026] MergeMix: A Unified Augmentation Paradigm for Visual and Multi-Modal Understanding☆23Feb 27, 2026Updated 7 months ago
- AFPQ code implementation☆23Nov 6, 2023Updated 2 years ago
- ☆50Jun 30, 2026Updated 3 months ago
- Code implementation of GPTAQ (https://arxiv.org/abs/2504.02692)☆96Jul 28, 2025Updated last year
- ☆35Dec 22, 2025Updated 9 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆111May 27, 2026Updated 4 months ago
- [NeurIPS 2025] Speculate Deep and Accurate☆25Aug 10, 2026Updated last month
- Dysl-vla:Official code for "DySL-VLA: Efficient Vision-Language-Action Model Inference via Dynamic-Static Layer-Skipping for Robot Manipu…☆26Jul 9, 2026Updated 2 months ago
- [CVPR 2026] OmniZip: Audio-Guided Dynamic Token Compression for Fast Omnimodal Large Language Models☆109Apr 20, 2026Updated 5 months ago
- Unofficial OpenPI extension experiment to build a more complete OpenPI-style VLA engineering stack: pi0.5 semantics, RTC, pi0.6 RECAP/MEM…☆41Jul 8, 2026Updated 2 months ago
- [ICML'25][TPAMI'26] Official implementation of paper "SparseVLM" and "SparseVLM+".☆280Jul 30, 2026Updated 2 months ago
- Official implementation of "Modeling Multi-Task Model Merging as Adaptive Projective Gradient Descent".☆23May 23, 2025Updated last year