[CVPR25 Highlight] A ChatGPT-Prompted Visual hallucination Evaluation Dataset, featuring over 100,000 data samples and four advanced evaluation modes. The dataset includes extensive contextual descriptions, counterintuitive images, and clear indicators of hallucination items.
β32Apr 16, 2025Updated last year
Alternatives and similar repositories for PhD
Users that are interested in PhD are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π curated list of awesome LMM hallucinations papers, methods & resources.β150Mar 23, 2024Updated 2 years ago
- Code for Reducing Hallucinations in Vision-Language Models via Latent Space Steeringβ119Nov 23, 2024Updated last year
- [CVPR 2024] TeachCLIP for Text-to-Video Retrievalβ42May 7, 2025Updated last year
- [ACL 2025] Cracking the Code of Hallucination in LVLMs with Vision-aware Head Divergenceβ20Jun 10, 2025Updated last year
- [ICCV 2025] ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Modelsβ52Jul 7, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Code for paper: Visual Signal Enhancement for Object Hallucination Mitigation in Multimodal Large language Modelsβ62Dec 18, 2024Updated last year
- [CVPR 2025] Devils in Middle Layers of Large Vision-Language Models: Interpreting, Detecting and Mitigating Object Hallucinations via Attβ¦β84Oct 9, 2025Updated 11 months ago
- [CVPR 2025] PyTorch implementation of Diff-IIβ30Aug 22, 2026Updated 3 weeks ago
- Official repository for Robust Multimodal Large Language Models Against Modality Conflictβ22Jul 9, 2025Updated last year
- [CVPR 2025] Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attentionβ69Jul 16, 2024Updated 2 years ago
- [ACL 2025 Findings] Official pytorch implementation of "Don't Miss the Forest for the Trees: Attentional Vision Calibration for Large Visβ¦β25Jul 21, 2024Updated 2 years ago
- Code for paper: Nullu: Mitigating Object Hallucinations in Large Vision-Language Models via HalluSpace Projectionβ64Mar 13, 2025Updated last year
- [ICCV 2025] Official repository of "Mitigating Object Hallucinations via Sentence-Level Early Intervention".β32Jul 2, 2026Updated 2 months ago
- [ACL 2026] WikiSeeker: Rethinking the Role of Vision-Language Models in Knowledge-Based Visual Question Answering.β15Jul 25, 2026Updated last month
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [CVPR'24] HallusionBench: You See What You Think? Or You Think What You See? An Image-Context Reasoning Benchmark Challenging for GPT-4V(β¦β342Oct 14, 2025Updated 10 months ago
- [NeurIPS 2025] More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Modelsβ82May 31, 2025Updated last year
- Repo for the paper "Words or Vision: Do Vision-Language Models Have Blind Faith in Text?" (CVPR 2025)β18Mar 31, 2026Updated 5 months ago
- This repo contains the code for the paper "Understanding and Mitigating Hallucinations in Large Vision-Language Models via Modular Attribβ¦β39Jul 14, 2025Updated last year
- [CVPR2025] Code Release of Patch Matters: Training-free Fine-grained Image Caption Enhancement via Local Perceptionβ25Jun 17, 2025Updated last year
- Official resource for paper Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models (ACL 20β¦β18Aug 12, 2024Updated 2 years ago
- [CVPR 2025 (Oral)] Mitigating Hallucinations in Large Vision-Language Models via DPO: On-Policy Data Hold the Keyβ114Jan 9, 2026Updated 8 months ago
- PyTorch Implementation of "Divide, Conquer and Combine: A Training-Free Framework for High-Resolution Image Perception in Multimodal Largβ¦β53Mar 2, 2026Updated 6 months ago
- Source code for EMNLP2022 paper "Finding Skill Neurons in Pre-trained Transformers via Prompt Tuning".β19Mar 13, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits β’ AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ICLR '25] Official Pytorch implementation of "Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations"β105Nov 30, 2025Updated 9 months ago
- Official repo for ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Modelsβ29Mar 24, 2025Updated last year
- β87Jul 28, 2025Updated last year
- [AAAI 2025] ConVis: Contrastive Decoding with Hallucination Visualization for Mitigating Hallucinations in Multimodal Large Language Modeβ¦β25Sep 26, 2024Updated last year
- [CVPR 2026] STAMP: Better, Stronger, Faster: Tackling the Trilemma in MLLM-based Segmentation with Simultaneous Textual Mask Predictionβ43Feb 21, 2026Updated 6 months ago
- [ICLR 2025] MLLM can see? Dynamic Correction Decoding for Hallucination Mitigationβ147Sep 11, 2025Updated last year
- π A curated list of resources dedicated to hallucination of multimodal large language models (MLLM).β1,042Sep 27, 2025Updated 11 months ago
- [ICLR 2026 Oral] πHallucination Begins Where Saliency Dropsβ69Feb 12, 2026Updated 7 months ago
- [COLM 2025] JailDAM: Jailbreak Detection with Adaptive Memory for Vision-Language Modelβ26Nov 25, 2025Updated 9 months ago
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Pytorch implementation of Detectiveβ14Jul 11, 2024Updated 2 years ago
- (NeXD @ CVPR 2025) Why We Feel: Breaking Boundaries in Emotional Reasoning with Multimodal Large Language Modelsβ33Sep 30, 2025Updated 11 months ago
- [ICLR 2026] Empowering Small VLMs to Think with Dynamic Memorization and Explorationβ19Mar 18, 2026Updated 5 months ago
- [CVPR2025] Official Repository for IMMUNE: Improving Safety Against Jailbreaks in Multi-modal LLMs via Inference-Time Alignmentβ29Jun 11, 2025Updated last year
- [CVPR'24 Highlight] Implementation of "Causal-CoG: A Causal-Effect Look at Context Generation for Boosting Multi-modal Language Models"β17Sep 12, 2024Updated 2 years ago
- This repo contains script to download MUSIC dataset from youtubeβ13Jan 19, 2024Updated 2 years ago
- Codes and Data for ICLR 2026 paper "LogicReward: Incentivizing LLM Reasoning via Step-Wise Logical Supervision"β25Jul 16, 2026Updated last month