[CVPR25 Highlight] A ChatGPT-Prompted Visual hallucination Evaluation Dataset, featuring over 100,000 data samples and four advanced evaluation modes. The dataset includes extensive contextual descriptions, counterintuitive images, and clear indicators of hallucination items.
β32Apr 16, 2025Updated last year
Alternatives and similar repositories for PhD
Users that are interested in PhD are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV 2023] GeoFormer for Homography Estimationβ35Dec 25, 2023Updated 2 years ago
- π curated list of awesome LMM hallucinations papers, methods & resources.β150Mar 23, 2024Updated 2 years ago
- [CVPR 2025] PyTorch implementation of Diff-IIβ28Feb 27, 2025Updated last year
- Code for Reducing Hallucinations in Vision-Language Models via Latent Space Steeringβ117Nov 23, 2024Updated last year
- [ACL 2025] Cracking the Code of Hallucination in LVLMs with Vision-aware Head Divergenceβ21Jun 10, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- In OLHWDB ,you can find the ptts files, this code can help you get the information of the pttsβ11Mar 8, 2022Updated 4 years ago
- Official implementation of "VideoSketcher: Video Models Prior Enable Versatile Sequential Sketch Generation"β15Apr 7, 2026Updated 3 months ago
- Source code for EMNLP2022 paper "Finding Skill Neurons in Pre-trained Transformers via Prompt Tuning".β18Mar 13, 2023Updated 3 years ago
- [ICLR '25] Official Pytorch implementation of "Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations"β106Nov 30, 2025Updated 7 months ago
- Official repo for ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Modelsβ28Mar 24, 2025Updated last year
- The first attempt to replicate o3-like visual clue-tracking reasoning capabilities.β64Jul 8, 2025Updated last year
- Code for paper: Nullu: Mitigating Object Hallucinations in Large Vision-Language Models via HalluSpace Projectionβ63Mar 13, 2025Updated last year
- Pytorch implementation of Detectiveβ13Jul 11, 2024Updated 2 years ago
- [ICLR 2026] Empowering Small VLMs to Think with Dynamic Memorization and Explorationβ18Mar 18, 2026Updated 4 months ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- This repo contains script to download MUSIC dataset from youtubeβ12Jan 19, 2024Updated 2 years ago
- Codes and Data for ICLR 2026 paper "LogicReward: Incentivizing LLM Reasoning via Step-Wise Logical Supervision"β23Updated this week
- VoCoT: Unleashing Visually Grounded Multi-Step Reasoning in Large Multi-Modal Modelsβ79Jul 13, 2024Updated 2 years ago
- [CVPR 2025 (Oral)] Mitigating Hallucinations in Large Vision-Language Models via DPO: On-Policy Data Hold the Keyβ111Jan 9, 2026Updated 6 months ago
- π A curated list of resources dedicated to hallucination of multimodal large language models (MLLM).β1,033Sep 27, 2025Updated 9 months ago
- TIFS2022: Decision-based Adversarial Attack with Frequency Mixupβ22Aug 8, 2023Updated 2 years ago
- β10Oct 21, 2024Updated last year
- [ICCV 2025] Official repository of "Mitigating Object Hallucinations via Sentence-Level Early Intervention".β31Jul 2, 2026Updated 2 weeks ago
- [CVPR 2025] VASparse: Towards Efficient Visual Hallucination Mitigation via Visual-Aware Token Sparsificationβ50Mar 24, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- β10Jan 19, 2022Updated 4 years ago
- Code and data for EMNLP 2023 paper "Grounding Visual Illusions in Language: Do Vision-Language Models Perceive Illusions Like Humans?"β15Jan 25, 2024Updated 2 years ago
- [ICLR 2025] MLLM can see? Dynamic Correction Decoding for Hallucination Mitigationβ146Sep 11, 2025Updated 10 months ago
- Project Page For "Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement"β635Jan 17, 2026Updated 6 months ago
- π΅οΈ ArXiv Agent v1.0 - Your Intelligent Research Assistantβ27Dec 29, 2025Updated 6 months ago
- This is the dataset for the competition "Clinical Brain Computer Interfaces Challenge" to be held at WCCI 2020 at Glasgow. There are the β¦β12Jan 20, 2022Updated 4 years ago
- This repo contains the code for the paper "Understanding and Mitigating Hallucinations in Large Vision-Language Models via Modular Attribβ¦β39Jul 14, 2025Updated last year
- β13Dec 28, 2023Updated 2 years ago
- Competition of Mechanisms: Tracing How Language Models Handle Facts and Counterfactuals; ACL 2024β13May 24, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [CVPR 2025] Devils in Middle Layers of Large Vision-Language Models: Interpreting, Detecting and Mitigating Object Hallucinations via Attβ¦β84Oct 9, 2025Updated 9 months ago
- [ICCV 2025] VisRL: Intention-Driven Visual Perception via Reinforced Reasoningβ47Nov 8, 2025Updated 8 months ago
- [CVPR'24] HallusionBench: You See What You Think? Or You Think What You See? An Image-Context Reasoning Benchmark Challenging for GPT-4V(β¦β342Oct 14, 2025Updated 9 months ago
- We're Not Using Videos Effectively (TMLR 2024)β17Feb 4, 2024Updated 2 years ago
- β¨β¨The Curse of Multi-Modalities (CMM): Evaluating Hallucinations of Large Multimodal Models across Language, Visual, and Audioβ54Jul 11, 2025Updated last year
- [ICLR'26] Traceable Evidence Enhanced Visual Grounded Reasoning: Evaluation and Methodologyβ92Jan 26, 2026Updated 5 months ago
- β86Jul 28, 2025Updated 11 months ago