[CVPR 2025 Workshop] PAINT (Paying Attention to INformed Tokens) is a plug-and-play framework that intervenes in the self-attention of the LLM and selectively boost the visual attention informed tokens to mitigate hallucination of Vision Language Models
☆20Jul 19, 2026Updated 2 months ago
Alternatives and similar repositories for PAINT
Users that are interested in PAINT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [AAAI 2025] HiRED strategically drops visual tokens in the image encoding stage to improve inference efficiency for High-Resolution Visio…☆58Apr 18, 2025Updated last year
- [ICLR 2025] Code for Self-Correcting Decoding with Generative Feedback for Mitigating Hallucinations in Large Vision-Language Models☆26Apr 14, 2025Updated last year
- ☆18Aug 1, 2024Updated 2 years ago
- Official pytorch implementation of "RITUAL: Random Image Transformations as a Universal Anti-hallucination Lever in Large Vision Language…☆14Dec 16, 2024Updated last year
- Mitigating Open-Vocabulary Caption Hallucinations (EMNLP 2024)☆19Oct 18, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆15Dec 16, 2025Updated 9 months ago
- [NeurIPS 2025] Poison as Cure: Visual Noise for Mitigating Object Hallucinations in LVMs☆37Sep 21, 2025Updated last year
- YAI 11 x @POZAlabs : Improving & Evaluating Music Generation with ComMU☆13Apr 5, 2023Updated 3 years ago
- [AAAI 2025] ConVis: Contrastive Decoding with Hallucination Visualization for Mitigating Hallucinations in Multimodal Large Language Mode…☆26Sep 26, 2024Updated 2 years ago
- ☆88Jul 28, 2025Updated last year
- Code for paper: Visual Signal Enhancement for Object Hallucination Mitigation in Multimodal Large language Models☆62Dec 18, 2024Updated last year
- ☆11Dec 14, 2022Updated 3 years ago
- CVPR2023: Few-Shot Learning with Visual Distribution Calibration and Cross-Modal Distribution Alignment☆14May 19, 2023Updated 3 years ago
- [CVPR 2025] Devils in Middle Layers of Large Vision-Language Models: Interpreting, Detecting and Mitigating Object Hallucinations via Att…☆84Oct 9, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Code for ICLR 2025 Paper: Visual Description Grounding Reduces Hallucinations and Boosts Reasoning in LVLMs☆25May 7, 2025Updated last year
- ☆12Nov 16, 2020Updated 5 years ago
- [ICLR 2025] MLLM can see? Dynamic Correction Decoding for Hallucination Mitigation☆147Sep 11, 2025Updated last year
- This repository contains the code of our paper 'Skip \n: A simple method to reduce hallucination in Large Vision-Language Models'.☆15Feb 12, 2024Updated 2 years ago
- ERGO (Efficient Reasoning & Guided Observation) is a large vision-language model trained with reinforcement learning on efficiency object…☆19Feb 25, 2026Updated 7 months ago
- Concept Learning Dynamics☆18Oct 29, 2024Updated last year
- A variant of Ahash written in C++.☆10Mar 20, 2023Updated 3 years ago
- ☆68May 19, 2025Updated last year
- Official code for **Prune Redundancy, Preserve Essence: Vision Token Compression in VLMs via Synergistic Importance-Diversity** (PruneSI…☆14Mar 25, 2026Updated 6 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆12May 30, 2024Updated 2 years ago
- To mitigate position bias in LLMs, especially in long-context scenarios, we scale only one dimension of LLMs, reducing position bias and …☆12Jun 18, 2024Updated 2 years ago
- Official code for paper "Reasoning Fails Where Step Flow Breaks" (ACL 2026)☆19Apr 19, 2026Updated 5 months ago
- The code and datasets of our ACM MM 2024 paper "Hallu-PI: Evaluating Hallucination in Multi-modal Large Language Models within Perturbed …☆11Sep 27, 2024Updated 2 years ago
- [NeurIPS 2025] More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models☆84May 31, 2025Updated last year
- [ACL 2025 Findings] Official pytorch implementation of "Don't Miss the Forest for the Trees: Attentional Vision Calibration for Large Vis…☆26Jul 21, 2024Updated 2 years ago
- Official implementation of the ECCV2024 paper: Generalizable Facial Expression Recognition☆22Sep 20, 2024Updated 2 years ago
- [AAAI‘24] The official PyTorch implimentation of our AAAI 2024 paper: Personalized LoRA for Human-Centered Text Understanding☆14Dec 11, 2023Updated 2 years ago
- dawei.li SemEval-2019 task3 EmoContext: Multi-Step Ensemble Neural Network for Sentiment Analysis in Textual Conversation☆16Jun 6, 2019Updated 7 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Template repo for Python projects, especially those focusing on machine learning and/or deep learning.☆15Jan 14, 2026Updated 8 months ago
- The source code and manually annotated datasets for our paper "Joint Multimodal Sentiment Analysis Based on Information Relevance"☆11Dec 17, 2022Updated 3 years ago
- Java-like Language with Static Information Flow Types☆16May 5, 2025Updated last year
- ☆15May 6, 2021Updated 5 years ago
- Crop growth stage modeling and classification☆12Apr 11, 2019Updated 7 years ago
- A novel graph-temporal model designed for subject-invariant and session-invariant motor imagery EEG (MI-EEG) classification☆12May 24, 2025Updated last year
- Code for "CLIP Behaves like a Bag-of-Words Model Cross-modally but not Uni-modally"☆29Aug 3, 2026Updated 2 months ago