The official repository of paper "Evaluating MLLMs with Multimodal Multi-image Reasoning Benchmark"
☆19Jun 20, 2025Updated last year
Alternatives and similar repositories for MMRB
Users that are interested in MMRB are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Spatial Aptitude Training for Multimodal Langauge Models☆34Feb 8, 2026Updated 8 months ago
- Official repository of IDEA-Bench☆41Jan 24, 2025Updated last year
- SB-Bench: Stereotype Bias Benchmark for Large Multimodal Models☆15Jun 26, 2026Updated 3 months ago
- A practical bilingual guide to staying safe and prepared at conferences in Brazil / 巴西参会实用攻略与自救指南☆17Apr 22, 2026Updated 5 months ago
- [ICLR'25 Oral] MMIE: Massive Multimodal Interleaved Comprehension Benchmark for Large Vision-Language Models☆35Nov 3, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official Code for CVPR 2024 paper: Permutation Equivariance of Transformers and Its Applications.☆19Nov 12, 2024Updated last year
- Code for CVPR 2024 Oral "Neural Lineage"☆17Jun 18, 2024Updated 2 years ago
- Official repository for “Reasoning in the Dark: Interleaved Vision-Text Reasoning in Latent Space”☆18Jan 27, 2026Updated 8 months ago
- [ECCV26]CoCo: Code as CoT for Text-to-Image Preview and Rare Concept Generation☆56Aug 7, 2026Updated 2 months ago
- Official Code for GPIC: A Giant Permissive Image Corpus for Visual Generation☆55Jun 4, 2026Updated 4 months ago
- Learning from Indirect Observations☆11Jul 16, 2021Updated 5 years ago
- ☆16Nov 12, 2024Updated last year
- This is a repository for awesome any2any work collection.☆32Oct 3, 2026Updated last week
- code for EMNLP 2024 paper: How do Large Language Models Learn In-Context? Query and Key Matrices of In-Context Heads are Two Towers for M…☆13Nov 17, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [ICCV 2025] Official code for paper: Beyond Text-Visual Attention: Exploiting Visual Cues for Effective Token Pruning in VLMs☆85Jul 1, 2025Updated last year
- ☆16Dec 25, 2025Updated 9 months ago
- ☆22Jan 22, 2026Updated 8 months ago
- [ACL2025 Findings] Migician: Revealing the Magic of Free-Form Multi-Image Grounding in Multimodal Large Language Models☆92May 20, 2025Updated last year
- In the context of Deep Learning: What is the right way to conduct example weighting? How do you understand loss functions and so-called …☆10Mar 4, 2021Updated 5 years ago
- [WACV 2024] Instruct Me More! Random Prompting for Visual In-Context Learning☆18May 7, 2025Updated last year
- Satellite package for LBP-TOP based face anti-spoofing☆11Oct 2, 2014Updated 12 years ago
- ☆19Feb 20, 2024Updated 2 years ago
- The VeriNet toolkit for verification of neural networks☆22Jul 2, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Official implementation for the paper "Video-Based Reward Modeling for Computer-Use Agents"☆17Mar 14, 2026Updated 6 months ago
- converting the pretrained tensorflow SoundNet model to pytorch☆14Jun 15, 2022Updated 4 years ago
- [ACL2026 oral] Uni-MMMU : A Massive Multi-discipline Multimodal Unified Benchmark☆27Apr 13, 2026Updated 5 months ago
- A simple PyTorch implementation of Learning Instance Activation Maps for Weakly Supervised Instance Segmentation, in CVPR 2019☆11Jun 18, 2020Updated 6 years ago
- ☆15Jul 30, 2020Updated 6 years ago
- Exploiting Class Activation Value for Partial-Label Learning, ICLR 2022 (poster)☆15Apr 18, 2022Updated 4 years ago
- ☆15Apr 3, 2023Updated 3 years ago
- codes for AAAI papar : Infrared-Visible Cross-Modal Person Re-Identification with an X Modality☆19Apr 8, 2020Updated 6 years ago
- The official implementation of the paper SAEdit: Token-level control for continuous image editing via Sparse AutoEncoder☆26Oct 19, 2025Updated 11 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- MACER: MAximizing CErtified Radius (ICLR 2020)☆31Jan 5, 2020Updated 6 years ago
- [EMNLP'24 (Main)] DRPO(Dynamic Rewarding with Prompt Optimization) is a tuning-free approach for self-alignment. DRPO leverages a search-…☆24Nov 17, 2024Updated last year
- ☆21Oct 1, 2024Updated 2 years ago
- [ICLR 2026] The official repository for paper "ThinkMorph: Emergent Properties in Multimodal Interleaved Chain-of-Thought Reasoning"☆200May 1, 2026Updated 5 months ago
- [AAAI 2025] Explore In-Context Segmentation via Latent Diffusion Models☆22Mar 25, 2025Updated last year
- [CVPR 2025] PACT: Pruning and Clustering-Based Token Reduction for Faster Visual Language Models☆60Jan 30, 2026Updated 8 months ago
- A Python toolkit for the OmniLabel benchmark providing code for evaluation and visualization☆23Feb 1, 2025Updated last year