π₯ [NeurIPS 2025] Official implementation of "Generate, but Verify: Reducing Visual Hallucination in Vision-Language Models with Retrospective Resampling (REVERSE)"
β58Jan 22, 2026Updated 6 months ago
Alternatives and similar repositories for reverse_vlm
Users that are interested in reverse_vlm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π₯ [ICML 2026] Official implementation of "Are LRMs Interruptible?"β19Jun 18, 2026Updated last month
- π₯ [ICLR 2025] Official PyTorch Model "Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark"β27Feb 9, 2025Updated last year
- Echo: "Constantly Improving Image Models Need Constantly Improving Benchmarks" (ICLR 2026)β20Jan 29, 2026Updated 6 months ago
- π₯ [CVPR 2024] Official implementation of "See, Say, and Segment: Teaching LMMs to Overcome False Premises (SESAME)"β47Jun 16, 2024Updated 2 years ago
- π₯π₯[NeurIPS2025]Exploring and mitigating semantic hallucinations in scene text perception and reasoningβ30Dec 11, 2025Updated 7 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A framework for pitting LLMs against each other in an evolving library of games ββ35Apr 20, 2025Updated last year
- Official Repository of VisGym: Diverse, Customizable, Scalable Environments for Multimodal Agentsβ114May 3, 2026Updated 3 months ago
- [ACL 2025] "CoT-UQ: Improving Response-wise Uncertainty Quantification in LLMs with Chain-of-Thought"β17Apr 3, 2025Updated last year
- [NeurIPS 2024] "Self-Calibrated Tuning of Vision-Language Models for Out-of-Distribution Detection"β13Oct 28, 2024Updated last year
- [ICML 2025] VistaDPO: Video Hierarchical Spatial-Temporal Direct Preference Optimization for Large Video Modelsβ42Jun 14, 2025Updated last year
- Recursive Visual Programming (ECCV 2024)β18Nov 20, 2024Updated last year
- β25Aug 2, 2024Updated 2 years ago
- [ICCV 2025] ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Modelsβ51Jul 7, 2025Updated last year
- β30Sep 2, 2025Updated 11 months ago
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- The official implementation of Preference Data Reward-Augmentation.β18May 1, 2025Updated last year
- [ACL 2025 Findings] Implicit Reasoning in Transformers is Reasoning through Shortcutsβ18Mar 11, 2025Updated last year
- β14Apr 23, 2025Updated last year
- code releaseβ38Jun 22, 2026Updated last month
- [ICLR 2025] MLLM can see? Dynamic Correction Decoding for Hallucination Mitigationβ148Sep 11, 2025Updated 10 months ago
- Code for ICLR 2025 Paper: Visual Description Grounding Reduces Hallucinations and Boosts Reasoning in LVLMsβ25May 7, 2025Updated last year
- β23Mar 31, 2026Updated 4 months ago
- π₯ [ICLR 2025] Official Benchmark Toolkits for "Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark"β45Nov 21, 2025Updated 8 months ago
- [NeurIPS 2025] Official Implementation for "Enhancing Vision-Language Model Reliability with Uncertainty-Guided Dropout Decoding"β22Dec 8, 2024Updated last year
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [ACL 2024] Mitigating Hallucinations in Large Vision-Language Models with Instruction Contrastive Decodingβ18Nov 10, 2025Updated 9 months ago
- Welcome to the official repository for Siren, a project aimed at understanding and mitigating harmful behaviors in large language models β¦β15Jun 14, 2026Updated last month
- [COLM 2025] "C3PO: Critical-Layer, Core-Expert, Collaborative Pathway Optimization for Test-Time Expert Re-Mixing"β21Apr 9, 2025Updated last year
- VHTestβ16Oct 31, 2024Updated last year
- Repo for the paper "Words or Vision: Do Vision-Language Models Have Blind Faith in Text?" (CVPR 2025)β18Mar 31, 2026Updated 4 months ago
- Code for Reducing Hallucinations in Vision-Language Models via Latent Space Steeringβ117Nov 23, 2024Updated last year
- Code and data for paper "Exploring Hallucination of Large Multimodal Models in Video Understanding: Benchmark, Analysis and Mitigation".β25Oct 22, 2025Updated 9 months ago
- [CVPR 2025] VASparse: Towards Efficient Visual Hallucination Mitigation via Visual-Aware Token Sparsificationβ50Mar 24, 2025Updated last year
- A tool to assist in the interpretation of learned features in sparse autoencoders (in particular the four SAE's trained by Joseph Bloom oβ¦β19Oct 4, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official repo for ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Modelsβ28Mar 24, 2025Updated last year
- (ICLR 2026)Official repository of 'ScaleCap: Inference-Time Scalable Image Captioning via Dual-Modality Debiasingββ60Jan 26, 2026Updated 6 months ago
- β15Apr 27, 2025Updated last year
- Code release for "UnSAMv2: Self-Supervised Learning Enables Segment Anything at Any Granularity"β84Feb 1, 2026Updated 6 months ago
- Code repository for the paper "The Inherent Limits of Pretrained LLMs: The Unexpected Convergence of Instruction Tuning and In-Context Leβ¦β14Jan 16, 2025Updated last year
- [CVPR 2025] Devils in Middle Layers of Large Vision-Language Models: Interpreting, Detecting and Mitigating Object Hallucinations via Attβ¦β84Oct 9, 2025Updated 10 months ago
- Official resource for paper Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models (ACL 20β¦β18Aug 12, 2024Updated last year