shengliu66 / VTI
View external linksLinks

Code for Reducing Hallucinations in Vision-Language Models via Latent Space Steering

☆103

Alternatives and similar repositories for VTI

Users that are interested in VTI are comparing it to the libraries listed below

Sorting:

Lackel / AGLA
View on GitHub
[CVPR 2025] Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention
☆61Jul 16, 2024Updated last year
TianyunYoung / Hallucination-Attribution
View on GitHub
This repo contains the code for the paper "Understanding and Mitigating Hallucinations in Large Vision-Language Models via Modular Attrib…
☆33Jul 14, 2025Updated 7 months ago
ustc-hyin / ClearSight
View on GitHub
Code for paper: Visual Signal Enhancement for Object Hallucination Mitigation in Multimodal Large language Models
☆52Dec 18, 2024Updated last year
nickjiang2378 / vlm-hallucinations
View on GitHub
[ICLR '25] Official Pytorch implementation of "Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations"
☆95Nov 30, 2025Updated 2 months ago
DAMO-NLP-SG / VCD
View on GitHub
[CVPR 2024 Highlight] Mitigating Object Hallucinations in Large Vision-Language Models through Visual Contrastive Decoding
☆378Oct 7, 2024Updated last year
NishilBalar / Awesome-LVLM-Hallucination
View on GitHub
up-to-date curated list of state-of-the-art Large vision language models hallucinations research work, papers & resources
☆265Feb 8, 2026Updated last week
LALBJ / PAI
View on GitHub
[ECCV 2024] Paying More Attention to Image: A Training-Free Method for Alleviating Hallucination in LVLMs
☆163Nov 6, 2024Updated last year
BillChan226 / HALC
View on GitHub
[ICML 2024] Official implementation for "HALC: Object Hallucination Reduction via Adaptive Focal-Contrast Decoding"
☆108Dec 4, 2024Updated last year
xing0047 / cca-llava
View on GitHub
[NeurIPS 2024] Mitigating Object Hallucination via Concentric Causal Attention
☆66Aug 30, 2025Updated 5 months ago
zjunlp / Deco
View on GitHub
[ICLR 2025] MLLM can see? Dynamic Correction Decoding for Hallucination Mitigation
☆133Sep 11, 2025Updated 5 months ago
sangminwoo / AvisC
View on GitHub
[ACL 2025 Findings] Official pytorch implementation of "Don't Miss the Forest for the Trees: Attentional Vision Calibration for Large Vis…
☆24Jul 21, 2024Updated last year
yuezih / less-is-more
View on GitHub
Less is More: Mitigating Multimodal Hallucination from an EOS Decision Perspective (ACL 2024)
☆57Oct 28, 2024Updated last year
VITA-Group / SEAL
View on GitHub
[COLM 2025] SEAL: Steerable Reasoning Calibration of Large Language Models for Free
☆52Apr 6, 2025Updated 10 months ago
sangminwoo / RITUAL
View on GitHub
Official pytorch implementation of "RITUAL: Random Image Transformations as a Universal Anti-hallucination Lever in Large Vision Language…
☆14Dec 16, 2024Updated last year
mrwu-mac / R-Bench
View on GitHub
[ICML2024] Repo for the paper `Evaluating and Analyzing Relationship Hallucinations in Large Vision-Language Models'
☆22Jan 1, 2025Updated last year
Ziwei-Zheng / Nullu
View on GitHub
Code for paper: Nullu: Mitigating Object Hallucinations in Large Vision-Language Models via HalluSpace Projection
☆50Mar 13, 2025Updated 11 months ago
1zhou-Wang / MemVR
View on GitHub
[ICML 2025] Official implementation of paper 'Look Twice Before You Answer: Memory-Space Visual Retracing for Hallucination Mitigation in…
☆171Sep 25, 2025Updated 4 months ago
LzVv123456 / VISTA
View on GitHub
☆71Jul 28, 2025Updated 6 months ago
THU-BPM / ICT
View on GitHub
Official repo for ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models
☆25Mar 24, 2025Updated 10 months ago
kaistAI / Volcano
View on GitHub
[NAACL 2024] Vision language model that reduces hallucinations through self-feedback guided revision. Visualizes attentions on image feat…
☆47Aug 21, 2024Updated last year
d-ailin / CLIP-Guided-Decoding
View on GitHub
☆17Aug 1, 2024Updated last year
mrwu-mac / ControlMLLM
View on GitHub
[NeurIPS2024] Repo for the paper `ControlMLLM: Training-Free Visual Prompt Learning for Multimodal Large Language Models'
☆204Jul 17, 2025Updated 6 months ago
Ziwei-Zheng / VaLSe
View on GitHub
A library of visualization tools for the interpretability and hallucination analysis of large vision-language models (LVLMs).
☆41May 22, 2025Updated 8 months ago
FreedomIntelligence / TRIM
View on GitHub
We introduce new approach, Token Reduction using CLIP Metric (TRIM), aimed at improving the efficiency of MLLMs without sacrificing their…
☆20Jan 11, 2026Updated last month
luka-group / vlm-knowledge-conflict
View on GitHub
Code for paper "Unraveling Cross-Modality Knowledge Conflicts in Large Vision-Language Models."
☆52Oct 19, 2024Updated last year
findalexli / mllm-dpo
View on GitHub
[ACL 2024] Multi-modal preference alignment remedies regression of visual instruction tuning on language model
☆47Nov 10, 2024Updated last year
mshukor / xl-vlms
View on GitHub
XL-VLMs: General Repository for eXplainable Large Vision Language Models
☆46Sep 8, 2025Updated 5 months ago
zhangce01 / DeGF
View on GitHub
[ICLR 2025] Code for Self-Correcting Decoding with Generative Feedback for Mitigating Hallucinations in Large Vision-Language Models
☆24Apr 14, 2025Updated 10 months ago
showlab / Awesome-MLLM-Hallucination
View on GitHub
📖 A curated list of resources dedicated to hallucination of multimodal large language models (MLLM).
☆979Sep 27, 2025Updated 4 months ago
ForJadeForest / LIVE-Learnable-In-Context-Vector
View on GitHub
【NeurIPS 2024】The implementation of LIVE: Learnable In-Context Vector for Visual Question Answering https://arxiv.org/abs/2406.13185
☆22May 31, 2025Updated 8 months ago
The-Martyr / CausalMM
View on GitHub
[ICLR 2025] Mitigating Modality Prior-Induced Hallucinations in Multimodal Large Language Models via Deciphering Attention Causality
☆60Jul 5, 2025Updated 7 months ago
mengchuang123 / VASparse-github
View on GitHub
[CVPR 2025] VASparse: Towards Efficient Visual Hallucination Mitigation via Visual-Aware Token Sparsification
☆49Mar 24, 2025Updated 10 months ago
DripNowhy / ETA
View on GitHub
[ICLR 2025] PyTorch Implementation of "ETA: Evaluating Then Aligning Safety of Vision Language Models at Inference Time"
☆29Jul 20, 2025Updated 6 months ago
paulgavrikov / vlm_shapebias
View on GitHub
Official code for "Can We Talk Models Into Seeing the World Differently?" (ICLR 2025).
☆27Jan 26, 2025Updated last year
DAMO-NLP-SG / CMM
View on GitHub
✨✨The Curse of Multi-Modalities (CMM): Evaluating Hallucinations of Large Multimodal Models across Language, Visual, and Audio
☆52Jul 11, 2025Updated 7 months ago
X-PLUG / mPLUG-HalOwl
View on GitHub
mPLUG-HalOwl: Multimodal Hallucination Evaluation and Mitigating
☆97Jan 29, 2024Updated 2 years ago
CaoYuanpu / BiPO
View on GitHub
Personalized Steering of Large Language Models: Versatile Steering Vectors Through Bi-directional Preference Optimization
☆42Jul 28, 2024Updated last year
jiazhen-code / PhD
View on GitHub
[CVPR25 Highlight] A ChatGPT-Prompted Visual hallucination Evaluation Dataset, featuring over 100,000 data samples and four advanced eval…
☆31Apr 16, 2025Updated 9 months ago
wjpoom / SPEC
View on GitHub
[CVPR 2024] The official implementation of paper "synthesize, diagnose, and optimize: towards fine-grained vision-language understanding"
☆50Jun 16, 2025Updated 7 months ago

shengliu66 / VTIView external linksLinks

Alternatives and similar repositories for VTI

shengliu66 / VTI
View external linksLinks