[NeurIPS 2025] AutoPrune, a general pruning method for LLM/VLM/VLA
☆20Oct 7, 2025Updated 9 months ago
Alternatives and similar repositories for AutoPrune
Users that are interested in AutoPrune are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR2026] The repo of paper "The Blind Spot of Adaptation: Quantifying and Mitigating Forgetting in Fine-tuned Driving Models"☆26Apr 7, 2026Updated 3 months ago
- DUET-VLM: Dual stage Unified Efficient Token reduction for VLM Training and Inference☆25May 21, 2026Updated 2 months ago
- [ICCV 2025] Official code for paper: Beyond Text-Visual Attention: Exploiting Visual Cues for Effective Token Pruning in VLMs☆84Jul 1, 2025Updated last year
- [CVPR 2026] Variation-aware Vision Token Dropping for Faster Large Vision-Language Models☆34May 27, 2026Updated last month
- ☆19Jan 29, 2026Updated 5 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Official Implementation of Paper FOLDER (ICCV2025) and Turbo (ECCV2024)☆15Jun 27, 2025Updated last year
- ☆21Mar 11, 2026Updated 4 months ago
- PyTorch Implementation of “SMC++: Masked Learning of Unsupervised Video Semantic Compression", an extended version of ICCV 2023 paper "No…☆39Jan 11, 2026Updated 6 months ago
- [NeurIPS 2025] Official repository for “FlowCut: Rethinking Redundancy via Information Flow for Efficient Vision-Language Models”☆32Dec 9, 2025Updated 7 months ago
- [EMNLP 2025 main 🔥] Code for "Stop Looking for Important Tokens in Multimodal Language Models: Duplication Matters More"☆121Oct 12, 2025Updated 9 months ago
- [TCSVT 2023] RDO-PTQ: Rate-Distortion Optimized Post-Training Quantization for Learned Image Compression☆22Nov 1, 2023Updated 2 years ago
- [ECCV 2026] Official code of "DriveVA: Video action models are zero-shot drivers"☆33Jul 8, 2026Updated last week
- Official Implementation of the TCSVT2023 paper: "Towards Robust Neural Image Compression: Adversarial Attack and Finetuning"☆17Apr 26, 2024Updated 2 years ago
- Feedforward Post Compression for Dynamic Gaussian Splatting (for standard GS sequence)☆20Dec 23, 2025Updated 6 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆35Jun 3, 2025Updated last year
- ☆29Jul 25, 2025Updated 11 months ago
- [TMLR 2026] Survey: https://arxiv.org/pdf/2507.20198☆371May 29, 2026Updated last month
- Collects papers on autonomous driving E2E learning, VLM/VLA and Hybrid systems, with organized research branches and trends in these fiel…☆205May 27, 2026Updated last month
- ☆27Mar 1, 2025Updated last year
- [ACL-2026 Findings] Implementation for HiPrune, a training-free visual token pruning method for VLM acceleration.☆58Apr 29, 2026Updated 2 months ago
- [TCSVT] Official repository of the paper "A Glimpse to Compress: Dynamic Visual Token Pruning for Large Vision-Language Models"☆98Jun 12, 2026Updated last month
- ☆12Dec 7, 2024Updated last year
- [ECCV 2024] Learning Unified Reference Representation for Unsupervised Multi-class Anomaly Detection☆19Sep 1, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A collection of VLMs papers, blogs, and projects, with a focus on VLMs in Autonomous Driving and related reasoning techniques.☆11Nov 16, 2024Updated last year
- [ICLR 2026] Ego-Scene Interactive Modeling for Autonomous Driving☆19Mar 11, 2026Updated 4 months ago
- Authors' PyTorch implementation of lossy image compression methods based on hierarchical VAEs☆89Oct 18, 2024Updated last year
- An autonomous driving simulation system that leverages fisheye camera and other sensing technologies☆12Feb 21, 2025Updated last year
- Official code for "Computationally-Efficient Neural Image Compression with Shallow Decoders", ICCV 2023☆35Oct 15, 2024Updated last year
- ☆12Dec 15, 2024Updated last year
- [ICCV 2025] Multi-Granular Spatio-Temporal Token Merging for Training-Free Acceleration of Video LLMs☆61Feb 2, 2026Updated 5 months ago
- [EMNLP 2025 Main] Video Compression Commander: Plug-and-Play Inference Acceleration for Video Large Language Models☆127May 14, 2026Updated 2 months ago
- [NeurIPS 2025 🔥] Official implementation for "Don't Just Chase “Highlighted Tokens” in MLLMs: Revisiting Visual Holistic Context Retenti…☆66Mar 5, 2026Updated 4 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Official PyTorch Implementation of HSVA (NeurIPS'21)☆27Apr 19, 2023Updated 3 years ago
- A concise implementation of SimCSE☆16Aug 2, 2021Updated 4 years ago
- Proposed to leverage human driving data to learn a social preference model of human driving using inverse reinforcement learning and then…☆11May 27, 2023Updated 3 years ago
- [NeurIPS 2025] HoliTom: Holistic Token Merging for Fast Video Large Language Models☆84Oct 10, 2025Updated 9 months ago
- [ICML 2025] CoreMatching: Co-adaptive Sparse Inference Framework for Comprehensive Acceleration of Vision Language Model☆16May 27, 2025Updated last year
- ☆52May 15, 2026Updated 2 months ago
- Official repository for VisionZip (CVPR 2025)☆443Jul 21, 2025Updated last year