[EMNLP 2025 Main Conference] Mitigating Hallucinations in Vision-Language Models through Image-Guided Head Suppression
☆16Dec 26, 2025Updated 9 months ago
Alternatives and similar repositories for SPIN
Users that are interested in SPIN are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACL 2025] Cracking the Code of Hallucination in LVLMs with Vision-aware Head Divergence☆20Jun 10, 2025Updated last year
- used VHDL to implement and simulate the IDEA-algorithm (International Data Encryption Algorithm). We will test the hardware-oriented impl…☆12Jan 15, 2019Updated 7 years ago
- This repo contains the code for the paper "Understanding and Mitigating Hallucinations in Large Vision-Language Models via Modular Attrib…☆39Jul 14, 2025Updated last year
- ☆19Jun 6, 2025Updated last year
- [NeurIPS 2025] Official Implementation for "Enhancing Vision-Language Model Reliability with Uncertainty-Guided Dropout Decoding"☆22Dec 8, 2024Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆15Nov 7, 2024Updated last year
- [NAACL 2025 Findings] Mitigating Hallucinations in Large Vision-Language Models via Summary-Guided Decoding☆16Feb 12, 2026Updated 7 months ago
- ☆10Jun 14, 2022Updated 4 years ago
- 😎 Awesome papers on token redundancy reduction☆14Mar 12, 2025Updated last year
- ☆13May 13, 2022Updated 4 years ago
- What do CLIP Vision Transformers learn? Feature Visualization can show you!☆15Aug 29, 2024Updated 2 years ago
- Activation-Steered Compression☆18Jan 30, 2026Updated 8 months ago
- [CVPR 2025] Devils in Middle Layers of Large Vision-Language Models: Interpreting, Detecting and Mitigating Object Hallucinations via Att…☆84Oct 9, 2025Updated last year
- We introduce new approach, Token Reduction using CLIP Metric (TRIM), aimed at improving the efficiency of MLLMs without sacrificing their…☆23Jan 11, 2026Updated 8 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official repo for ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models☆31Mar 24, 2025Updated last year
- [ECCV 2024] CLAMP-ViT: Contrastive Data-Free Learning for Adaptive Post-Training Quantization of ViTs☆20Jul 2, 2024Updated 2 years ago
- (ICLR 2026) Unveiling Super Experts in Mixture-of-Experts Large Language Models☆44Sep 25, 2025Updated last year
- VHTest☆16Oct 31, 2024Updated last year
- [COLM'25] CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing☆19Jun 25, 2025Updated last year
- Implementation Code for "LLM-based Medical Assistant Personalization with Short- and Long-Term Memory Coordination"☆14May 17, 2026Updated 4 months ago
- Official Implementation of LANTERN (ICLR'25) and LANTERN++(ICLRW-SCOPE'25)☆22Mar 5, 2025Updated last year
- Code for boomerang distillation enables zero-shot model size interpolation.☆23Jul 10, 2026Updated 2 months ago
- Simple wgpu based SLAM map viewer.☆10May 5, 2020Updated 6 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- The code and data for "Summary-Oriented Vision Modeling for Multimodal Abstractive Summarization"☆11May 16, 2023Updated 3 years ago
- Official resource for paper Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models (ACL 20…☆19Aug 12, 2024Updated 2 years ago
- ☆17Feb 23, 2025Updated last year
- Code and data for ACL 2024 paper on 'Cross-Modal Projection in Multimodal LLMs Doesn't Really Project Visual Attributes to Textual Space'☆19Jul 21, 2024Updated 2 years ago
- This code is provided for reproducibility of results in the paper: Multiview Aerial Visual Recognition (MAVREC): Can Multi-view Improve A…☆24Feb 6, 2025Updated last year
- mstar: Optimizing memory architecture for every LLM task as executable Python code.☆17Jun 12, 2026Updated 3 months ago
- Learning Motion and Temporal Cues for Unsupervised Video Object Segmentation[TNNLS2024]☆14May 6, 2025Updated last year
- [ACL 2024] Mitigating Hallucinations in Large Vision-Language Models with Instruction Contrastive Decoding☆19Nov 10, 2025Updated 10 months ago
- [TPAMI 2026] Recurrence Meets Transformers for Universal Multimodal Retrieval☆15Dec 15, 2025Updated 9 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official Implementation of CODE☆18Sep 26, 2024Updated 2 years ago
- [CVPR 2025] Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention☆71Jul 16, 2024Updated 2 years ago
- repository for "Exploiting Proximity-Aware Tasks for Embodied Social Navigation" paper code☆12Nov 16, 2023Updated 2 years ago
- [EMNLP 2024] Preserving Multi-Modal Capabilities of Pre-trained VLMs for Improving Vision-Linguistic Compositionality☆24Oct 8, 2024Updated 2 years ago
- Repository for "Prompt-MIL: Boosting Multi-Instance Learning Schemes via Task-specific Prompt Tuning" (MICCAI2023)☆19Oct 23, 2023Updated 2 years ago
- [ICLR 2025] Data-Augmented Phrase-Level Alignment for Mitigating Object Hallucination☆21Jan 27, 2025Updated last year
- [ISCA 2025] Official Implementation of "MicroScopiQ: Accelerating Foundational Models through Outlier-Aware Microscaling Quantization"☆26Oct 30, 2025Updated 11 months ago