[ICLR 26] Visual Multi-Agent System: Mitigating Hallucination Snowballing via Visual Flow
☆46Oct 3, 2025Updated 10 months ago
Alternatives and similar repositories for ViF
Users that are interested in ViF are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- TokenAR: Multiple Subject Generation via Autoregressive Token-level enhancement☆22Aug 4, 2026Updated last week
- [CVPR 2026] Boosting Reasoning in Large Multimodal Models via Activation Replay☆24May 7, 2026Updated 3 months ago
- Official code repository for Med-CMR : "A Fine-Grained Benchmark Integrating Visual Evidence and Clinical Logic for Medical Complex Multi…☆26Dec 10, 2025Updated 8 months ago
- Official repository for "Human-MME: A Holistic Evaluation Benchmark for Human-Centric Multimodal Large Language Models"☆23Dec 2, 2025Updated 8 months ago
- ☆23May 26, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICML 2026] The official code of FeRA: Frequency–Energy Constrained Routing for Effective Diffusion Adaptation Fine-Tuning☆29Dec 27, 2025Updated 7 months ago
- From Large Angles to Consistent Faces: Identity-Preserving Video Generation via Mixture of Facial Experts☆28Jan 12, 2026Updated 6 months ago
- ☆30Nov 28, 2025Updated 8 months ago
- ☆92Feb 5, 2026Updated 6 months ago
- ☆183Jun 8, 2026Updated 2 months ago
- [ICML 2026] Transform Trained Transformer for Accelerating Native 4K Video Generation☆41Dec 16, 2025Updated 7 months ago
- Code for paper: Visual Signal Enhancement for Object Hallucination Mitigation in Multimodal Large language Models☆62Dec 18, 2024Updated last year
- CLEAR: Context-Aware Learning with End-to-End Mask-Free Inference for Adaptive Video Subtitle Removal☆23May 25, 2026Updated 2 months ago
- ☆29Mar 22, 2026Updated 4 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ICLR 2026 Oral] 🎉Hallucination Begins Where Saliency Drops☆66Feb 12, 2026Updated 5 months ago
- The official code of Refinement Provenance Inference: Detecting LLM-Refined Training Prompts from Model Behavior☆22Jan 6, 2026Updated 7 months ago
- Introduction about AWESOME_ENTROPY+LRM_PAPERS☆32Dec 16, 2025Updated 7 months ago
- ☆14Feb 24, 2026Updated 5 months ago
- Official repository for Robust Multimodal Large Language Models Against Modality Conflict☆22Jul 9, 2025Updated last year
- [ICLR 2025] Data-Augmented Phrase-Level Alignment for Mitigating Object Hallucination☆22Jan 27, 2025Updated last year
- Official code for paper "Reasoning Fails Where Step Flow Breaks" (ACL 2026)☆18Apr 19, 2026Updated 3 months ago
- ☆32Mar 17, 2026Updated 4 months ago
- ☆19Jun 6, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This repository is the official implementation of "Look-Back: Implicit Visual Re-focusing in MLLM Reasoning".☆99Jul 10, 2025Updated last year
- ☆16Jan 12, 2026Updated 6 months ago
- ☆15Apr 6, 2026Updated 4 months ago
- [CVPR2026] Official codebase for the paper "Reasoning Within the Mind: Dynamic Multimodal Interleaving in Latent Space"☆87May 12, 2026Updated 2 months ago
- [AAAI 2026] SIFThinker: Spatially-Aware Image Focus for Visual Reasoning☆22Dec 2, 2025Updated 8 months ago
- ☆42Jun 17, 2026Updated last month
- [MICCAI 2025] GL-LCM: Global-Local Latent Consistency Models for Fast High-Resolution Bone Suppression in Chest X-Ray Images☆17Mar 12, 2026Updated 4 months ago
- Official repo for "PAPO: Perception-Aware Policy Optimization for Multimodal Reasoning"☆156Feb 4, 2026Updated 6 months ago
- Official implementation of "FoundDiff: Foundational Diffusion Model for Generalizable Low-Dose CT Denoising"☆22Feb 7, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ACL 2025] Cracking the Code of Hallucination in LVLMs with Vision-aware Head Divergence☆22Jun 10, 2025Updated last year
- 4DThinker: Thinking with 4D Imagery for Dynamic Spatial Understanding☆78May 26, 2026Updated 2 months ago
- [ACMMM25] Crisp-sam2: Sam2 with cross-modal interaction and semantic prompting for multi-organ segmentation☆39Jul 6, 2025Updated last year
- ☆42Nov 12, 2025Updated 8 months ago
- ☆88Jul 28, 2025Updated last year
- MLLM hallucination, LVLM, LLM, Hallucination Mitigation, Training-free hallucination mitigation☆34Jul 13, 2026Updated 3 weeks ago
- ☆31Apr 9, 2026Updated 4 months ago