DUET-VLM: Dual stage Unified Efficient Token reduction for VLM Training and Inference
☆25May 21, 2026Updated 2 months ago
Alternatives and similar repositories for DUET-VLM
Users that are interested in DUET-VLM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- PixelPrune: Pixel-Level Adaptive Visual Token Reduction via Predictive Coding☆28Jun 10, 2026Updated last month
- [NeurIPS 2025] AutoPrune, a general pruning method for LLM/VLM/VLA☆20Oct 7, 2025Updated 9 months ago
- ☆15Apr 15, 2026Updated 3 months ago
- [NeurIPS 2025] Official repository for “FlowCut: Rethinking Redundancy via Information Flow for Efficient Vision-Language Models”☆32Dec 9, 2025Updated 7 months ago
- [CVPR 2025] DivPrune: Diversity-based Visual Token Pruning for Large Multimodal Models☆86Apr 16, 2026Updated 3 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Official code of Why Knowledge Distillation Works in Generative Models: A Minimal Working Explanation (NeurIPS 2025)☆17Dec 23, 2025Updated 6 months ago
- Dataset Distillation via Vision-Language Category Prototype (ICCV 2025)☆16Mar 20, 2026Updated 4 months ago
- "DeepKD: A Deeply Decoupled and Denoised Knowledge Distillation Trainer" [NeurIPS 2025 Accepted]☆19May 22, 2025Updated last year
- WorldCache: Content-Aware Caching for Accelerated Video World Models☆21Jun 28, 2026Updated 3 weeks ago
- (ICLR 2025 Spotlight) Official code repository for Interleaved Scene Graph.☆31Aug 7, 2025Updated 11 months ago
- ☆16Sep 29, 2024Updated last year
- Source codes for our paper "Neural Temporality Adaptation for Document Classification: Diachronic Word Embeddings and Domain Adaptation M…☆12Apr 20, 2021Updated 5 years ago
- A collection of VLMs papers, blogs, and projects, with a focus on VLMs in Autonomous Driving and related reasoning techniques.☆11Nov 16, 2024Updated last year
- Code for "Are “Hierarchical” Visual Representations Hierarchical?" in NeurIPS Workshop for Symmetry and Geometry in Neural Representation…☆23Nov 8, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- OnlineEWC and EWC++ implementations, the online versions of Elastic weight consolidation☆11Sep 4, 2019Updated 6 years ago
- Official implementation of "ApET: Approximation-Error Guided Token Compression for Efficient VLMs" (CVPR 2026)☆29Jun 29, 2026Updated 3 weeks ago
- ☆10Feb 4, 2025Updated last year
- ☆14May 9, 2023Updated 3 years ago
- ☆35Apr 16, 2026Updated 3 months ago
- This repo uses a combination of logits and feature distillation method to teach the PSPNet model of ResNet18 backbone with the PSPNet mod…☆11Sep 30, 2021Updated 4 years ago
- Zoom-Refine: Boosting High-Resolution Multimodal Understanding via Localized Zoom and Self-Refinement☆19Jul 4, 2026Updated 2 weeks ago
- Official code repo of PIN: Positional Insert Unlocks Object Localisation Abilities in VLMs☆26Jan 14, 2025Updated last year
- Towards Efficient Multimodal Large Language Models: A Survey on Token Compression☆212Jun 29, 2026Updated 3 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- 对youtu的training_free_grpo的测试以及修改☆24Nov 6, 2025Updated 8 months ago
- [ICLR 2026] Multi-Head Low-Rank Attention☆34Apr 16, 2026Updated 3 months ago
- [ECCV 2026] StAR: Segment Anything Reasoner☆25Apr 2, 2026Updated 3 months ago
- ☆15Feb 27, 2024Updated 2 years ago
- [EMNLP 2025 main 🔥] Code for "Stop Looking for Important Tokens in Multimodal Language Models: Duplication Matters More"☆121Oct 12, 2025Updated 9 months ago
- [WACV 2026] ZonUI-3B — A lightweight, resolution-aware GUI grounding model trained with only 24K samples on a single RTX 4090.☆26Jan 2, 2026Updated 6 months ago
- [CVPR'26 Findings] Source code for "RADSeg Unleashing Parameter and Compute Efficient Zero-Shot Open-Vocabulary Segmentation Using Agglom…☆60May 31, 2026Updated last month
- ☆16Nov 25, 2021Updated 4 years ago
- [ICML2022] "Identity-Disentangled Adversarial Augmentation for Self-Supervised Learning"☆10Jul 24, 2022Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- CVPR25☆28Jul 2, 2025Updated last year
- Simple Contourlet Transform in Python☆17Jul 14, 2021Updated 5 years ago
- HWFI: Hybrid Warping Fusion for Video Frame Interpolation. IJCV 2022☆11Sep 7, 2022Updated 3 years ago
- Official Pytorch Implementation of BPKD : Boundary Privileged Knowledge Distillation For Semantic Segmentation☆25Oct 25, 2023Updated 2 years ago
- A simple visual test-time scaling method for GUI agent grounding☆26Dec 7, 2025Updated 7 months ago
- A Practical Zoom-in GUI Grounding and Behavior-Based Evaluation method.☆25Dec 8, 2025Updated 7 months ago
- Implementation of Latent Replay, a Continual Learning strategy for Real-Time / On The Edge applications☆14May 7, 2020Updated 6 years ago