Repo for the paper "Words or Vision: Do Vision-Language Models Have Blind Faith in Text?" (CVPR 2025)
☆18Mar 31, 2026Updated 4 months ago
Alternatives and similar repositories for blind-faith-in-text
Users that are interested in blind-faith-in-text are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR 2026 Oral] FINER: MLLMs Hallucinate under Fine-grained Negative Queries☆18Jul 6, 2026Updated last month
- Official repository for Robust Multimodal Large Language Models Against Modality Conflict☆22Jul 9, 2025Updated last year
- [NeurIPS 2025] More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models☆82May 31, 2025Updated last year
- [CVPR 2025 Highlight] Official implementation of HySAC, a hyperbolic safety-aware vision-language model for safer multimodal retrieval an…☆31Apr 8, 2025Updated last year
- Ἀνατομή is a PyTorch library to analyze representation of neural networks☆13Jan 31, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [ICLR 2025] Data-Augmented Phrase-Level Alignment for Mitigating Object Hallucination☆21Jan 27, 2025Updated last year
- A Massive Multi-Discipline Lecture Understanding Benchmark☆34Apr 20, 2026Updated 4 months ago
- Multi-Stage Vision Token Dropping: Towards Efficient Multimodal Large Language Model☆38Jan 8, 2025Updated last year
- Incomplete Multi-view Clustering via Diffusion Contrastive Generation☆31Mar 22, 2026Updated 5 months ago
- Code and dataset for NAACL 2022 paper "CoSIm: Commonsense Reasoning for Counterfactual Scene Imagination" Hyounghun Kim, Abhay Zala, Mohi…☆16Nov 26, 2022Updated 3 years ago
- SHU Graduate Thesis Typst Template☆12Apr 25, 2025Updated last year
- [MICCAI 2024 - Early Accept] BiasPruner: Debiased Continual Learning for Medical Image Classification☆13Oct 30, 2024Updated last year
- The source code for "LaSe-E2V: Towards Language-guided Semantic-Aware Event-to-Video Reconstruction"☆10Jul 5, 2024Updated 2 years ago
- Code of ACM MM 2023 Paper: A Symbolic Characters Aware Model for Solving Geometry Problems☆16Dec 27, 2023Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- CVPR2025: Benchmarking Large Vision-Language Models via Directed Scene Graph for Comprehensive Image Captioning☆40Mar 21, 2025Updated last year
- Evaluation and dataset construction code for the CVPR 2025 paper "Vision-Language Models Do Not Understand Negation"☆49Feb 26, 2026Updated 6 months ago
- ☆10Sep 13, 2022Updated 3 years ago
- [CVPR25 Highlight] A ChatGPT-Prompted Visual hallucination Evaluation Dataset, featuring over 100,000 data samples and four advanced eval…☆32Apr 16, 2025Updated last year
- [ECCV 2026] VKnowU: Evaluating Visual Knowledge Understanding in Multimodal LLMs☆18Feb 3, 2026Updated 6 months ago
- [ECCV'24] Official Implementation of Autoregressive Visual Entity Recognizer.☆14Mar 2, 2024Updated 2 years ago
- Code for paper "Unraveling Cross-Modality Knowledge Conflicts in Large Vision-Language Models."☆55Oct 19, 2024Updated last year
- Self-supervised adversarial masking for point clouds☆11Jul 12, 2023Updated 3 years ago
- ☆45Mar 31, 2025Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Code for "Thinking Forward: Memory-Efficient Federated Finetuning of Language Models" (NeurIPS 2024). Spry is a federated learning al…☆13Oct 8, 2024Updated last year
- Code repo for KDD'22 paper : 'RES: A Robust Framework for Guiding Visual Explanation'☆32Aug 21, 2022Updated 4 years ago
- Contrastive Continual Learning with Importance Sampling and Prototype-Instance Relation Distillation☆12Jul 22, 2024Updated 2 years ago
- Official repository for “Reasoning in the Dark: Interleaved Vision-Text Reasoning in Latent Space”☆18Jan 27, 2026Updated 7 months ago
- [ICLR25] Official Implementation of "Decoupling Angles and Strength in Low-rank Adaptation"☆15Dec 12, 2025Updated 8 months ago
- [CVPR 2022] DiSparse: Disentangled Sparsification for Multitask Model Compression☆13Sep 6, 2022Updated 3 years ago
- ☆44Jan 12, 2026Updated 7 months ago
- ☆15Mar 18, 2026Updated 5 months ago
- ☆14Jul 8, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- PFLoRA-lib: Personalized Federated Learning with LoRA Algorithm Library focusing on privacy-protection, federated-learning, Citation, Ext…☆14Sep 19, 2024Updated last year
- [AAAI 2025] Official Implementation of I-HallA v1.0☆16Feb 2, 2025Updated last year
- ☆10Jan 31, 2022Updated 4 years ago
- ☆20Mar 19, 2025Updated last year
- ☆12Aug 9, 2022Updated 4 years ago
- Official codes of "Sketch-in-Latents: Eliciting Unified Reasoning in MLLMs"☆17Feb 15, 2026Updated 6 months ago
- ☆17Oct 22, 2024Updated last year