[EMNLP 2025 Main Conference] Mitigating Hallucinations in Vision-Language Models through Image-Guided Head Suppression
☆15Dec 26, 2025Updated 8 months ago
Alternatives and similar repositories for SPIN
Users that are interested in SPIN are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACL 2025] Cracking the Code of Hallucination in LVLMs with Vision-aware Head Divergence☆21Jun 10, 2025Updated last year
- used VHDL to implement and simulate the IDEA-algorithm (International Data Encryption Algorithm). We will test the hardware-oriented impl…☆12Jan 15, 2019Updated 7 years ago
- This repo contains the code for the paper "Understanding and Mitigating Hallucinations in Large Vision-Language Models via Modular Attrib…☆39Jul 14, 2025Updated last year
- ☆19Jun 6, 2025Updated last year
- [NeurIPS 2025] Official Implementation for "Enhancing Vision-Language Model Reliability with Uncertainty-Guided Dropout Decoding"☆22Dec 8, 2024Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆15Nov 7, 2024Updated last year
- ☆10Jun 14, 2022Updated 4 years ago
- 😎 Awesome papers on token redundancy reduction☆14Mar 12, 2025Updated last year
- What do CLIP Vision Transformers learn? Feature Visualization can show you!☆15Aug 29, 2024Updated 2 years ago
- Accelerator RTL inspired by VEGETA [HPCA'23] and MicroScopiQ [ISCA'25]☆16Nov 11, 2025Updated 9 months ago
- [CVPR 2025] Devils in Middle Layers of Large Vision-Language Models: Interpreting, Detecting and Mitigating Object Hallucinations via Att…☆85Oct 9, 2025Updated 10 months ago
- We introduce new approach, Token Reduction using CLIP Metric (TRIM), aimed at improving the efficiency of MLLMs without sacrificing their…☆22Jan 11, 2026Updated 7 months ago
- Official repo for ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models☆28Mar 24, 2025Updated last year
- (ICLR 2026) Unveiling Super Experts in Mixture-of-Experts Large Language Models☆44Sep 25, 2025Updated 11 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- [ECCV 2024] CLAMP-ViT: Contrastive Data-Free Learning for Adaptive Post-Training Quantization of ViTs☆20Jul 2, 2024Updated 2 years ago
- Distillation Self-Knowledge From Contrastive Links to Classify Graph Nodes Without Passing Messages.☆15Jun 17, 2021Updated 5 years ago
- VHTest☆16Oct 31, 2024Updated last year
- ☆29Mar 12, 2026Updated 5 months ago
- [ECCV 2024] Paying More Attention to Image: A Training-Free Method for Alleviating Hallucination in LVLMs☆172Nov 6, 2024Updated last year
- Simple wgpu based SLAM map viewer.☆10May 5, 2020Updated 6 years ago
- Code for boomerang distillation enables zero-shot model size interpolation.☆22Jul 10, 2026Updated last month
- [ICLR 2025] Drop-Upcycling: Training Sparse Mixture of Experts with Partial Re-initialization☆25Oct 5, 2025Updated 10 months ago
- ☆15Aug 20, 2026Updated last week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- The code and data for "Summary-Oriented Vision Modeling for Multimodal Abstractive Summarization"☆11May 16, 2023Updated 3 years ago
- Official resource for paper Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models (ACL 20…☆18Aug 12, 2024Updated 2 years ago
- Code and data for ACL 2024 paper on 'Cross-Modal Projection in Multimodal LLMs Doesn't Really Project Visual Attributes to Textual Space'☆18Jul 21, 2024Updated 2 years ago
- ☆17Feb 23, 2025Updated last year
- [ICLR 2025] FLAT: LLM Unlearning via Loss Adjustment with Only Forget Data☆14Feb 26, 2025Updated last year
- mstar: Optimizing memory architecture for every LLM task as executable Python code.☆16Jun 12, 2026Updated 2 months ago
- [ACL 2024] Mitigating Hallucinations in Large Vision-Language Models with Instruction Contrastive Decoding☆18Nov 10, 2025Updated 9 months ago
- Recurrence Meets Transformers for Universal Multimodal Retrieval☆15Dec 15, 2025Updated 8 months ago
- Official Implementation of CODE☆17Sep 26, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Reading list for multimodal sequence learning☆14Sep 4, 2023Updated 2 years ago
- [CVPR 2025] Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention☆69Jul 16, 2024Updated 2 years ago
- repository for "Exploiting Proximity-Aware Tasks for Embodied Social Navigation" paper code☆12Nov 16, 2023Updated 2 years ago
- TypeScript agents for real applications.☆26Aug 5, 2026Updated 3 weeks ago
- [EMNLP 2024] Preserving Multi-Modal Capabilities of Pre-trained VLMs for Improving Vision-Linguistic Compositionality☆24Oct 8, 2024Updated last year
- Repository for "Prompt-MIL: Boosting Multi-Instance Learning Schemes via Task-specific Prompt Tuning" (MICCAI2023)☆19Oct 23, 2023Updated 2 years ago
- [ICLR 2025] Data-Augmented Phrase-Level Alignment for Mitigating Object Hallucination☆21Jan 27, 2025Updated last year