[ICLR 26] Visual Multi-Agent System: Mitigating Hallucination Snowballing via Visual Flow
☆46Oct 3, 2025Updated 10 months ago
Alternatives and similar repositories for ViF
Users that are interested in ViF are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆18Jul 31, 2025Updated last year
- TokenAR: Multiple Subject Generation via Autoregressive Token-level enhancement☆21Aug 4, 2026Updated 3 weeks ago
- [CVPR 2026] Boosting Reasoning in Large Multimodal Models via Activation Replay☆23Updated this week
- Official code repository for Med-CMR : "A Fine-Grained Benchmark Integrating Visual Evidence and Clinical Logic for Medical Complex Multi…☆26Dec 10, 2025Updated 8 months ago
- Official repository for "Human-MME: A Holistic Evaluation Benchmark for Human-Centric Multimodal Large Language Models"☆22Dec 2, 2025Updated 8 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆23May 26, 2025Updated last year
- [ICML 2026] The official code of FeRA: Frequency–Energy Constrained Routing for Effective Diffusion Adaptation Fine-Tuning☆28Dec 27, 2025Updated 8 months ago
- From Large Angles to Consistent Faces: Identity-Preserving Video Generation via Mixture of Facial Experts☆27Jan 12, 2026Updated 7 months ago
- ☆29Nov 28, 2025Updated 9 months ago
- ☆92Feb 5, 2026Updated 6 months ago
- ☆184Jun 8, 2026Updated 2 months ago
- [ICML 2026] Transform Trained Transformer for Accelerating Native 4K Video Generation☆41Dec 16, 2025Updated 8 months ago
- Code for paper: Visual Signal Enhancement for Object Hallucination Mitigation in Multimodal Large language Models☆63Dec 18, 2024Updated last year
- [CVPR 2026] Soul: Breathe Life into Digital Human for High-fidelity Long-term Multimodal Animation☆64Dec 16, 2025Updated 8 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- CLEAR: Context-Aware Learning with End-to-End Mask-Free Inference for Adaptive Video Subtitle Removal☆24May 25, 2026Updated 3 months ago
- Official code for **Prune Redundancy, Preserve Essence: Vision Token Compression in VLMs via Synergistic Importance-Diversity** (PruneSI…☆13Mar 25, 2026Updated 5 months ago
- [ICCV 2025] Official repository of "Mitigating Object Hallucinations via Sentence-Level Early Intervention".☆32Jul 2, 2026Updated last month
- ☆31Jan 11, 2026Updated 7 months ago
- [ICLR 2026 Oral] 🎉Hallucination Begins Where Saliency Drops☆69Feb 12, 2026Updated 6 months ago
- The official code of Refinement Provenance Inference: Detecting LLM-Refined Training Prompts from Model Behavior☆22Jan 6, 2026Updated 7 months ago
- ☆14Feb 24, 2026Updated 6 months ago
- Official repository for Robust Multimodal Large Language Models Against Modality Conflict☆22Jul 9, 2025Updated last year
- [ICLR 2025] Data-Augmented Phrase-Level Alignment for Mitigating Object Hallucination☆21Jan 27, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official code for paper "Reasoning Fails Where Step Flow Breaks" (ACL 2026)☆19Apr 19, 2026Updated 4 months ago
- ☆32Mar 17, 2026Updated 5 months ago
- ☆19Jun 6, 2025Updated last year
- This repository is the official implementation of "Look-Back: Implicit Visual Re-focusing in MLLM Reasoning".☆98Jul 10, 2025Updated last year
- ☆15Jan 12, 2026Updated 7 months ago
- ☆15Apr 6, 2026Updated 4 months ago
- [CVPR2026] Official codebase for the paper "Reasoning Within the Mind: Dynamic Multimodal Interleaving in Latent Space"☆88May 12, 2026Updated 3 months ago
- [AAAI 2026] SIFThinker: Spatially-Aware Image Focus for Visual Reasoning☆23Dec 2, 2025Updated 8 months ago
- ☆18Jul 14, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [EMNLP'26 Findings] OPD-Evolver☆43Jun 17, 2026Updated 2 months ago
- [MICCAI 2025] GL-LCM: Global-Local Latent Consistency Models for Fast High-Resolution Bone Suppression in Chest X-Ray Images☆17Mar 12, 2026Updated 5 months ago
- Official repo for "PAPO: Perception-Aware Policy Optimization for Multimodal Reasoning"☆158Feb 4, 2026Updated 6 months ago
- Official implementation of "FoundDiff: Foundational Diffusion Model for Generalizable Low-Dose CT Denoising"☆22Feb 7, 2026Updated 6 months ago
- [ACMMM25] Crisp-sam2: Sam2 with cross-modal interaction and semantic prompting for multi-organ segmentation☆39Jul 6, 2025Updated last year
- ☆42Nov 12, 2025Updated 9 months ago
- ☆88Jul 28, 2025Updated last year