The official repository of paper "Evaluating MLLMs with Multimodal Multi-image Reasoning Benchmark"
☆19Jun 20, 2025Updated last year
Alternatives and similar repositories for MMRB
Users that are interested in MMRB are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆19Jun 22, 2024Updated 2 years ago
- Dataset for the investigation of visual semiotics, and how specific visual features and design choices can elicit specific emotions, thou…☆11Dec 13, 2023Updated 2 years ago
- Spatial Aptitude Training for Multimodal Langauge Models☆33Feb 8, 2026Updated 6 months ago
- Reproduction code for paper "MineExplorer: Evaluating Open-World Exploration of MLLM Agents in Minecraft"☆21Jun 12, 2026Updated 2 months ago
- Code to implement the experiments in "Post-training Quantization for Neural Networks with Provable Guarantees" by Jinjie Zhang, Yixuan Zh…☆10Jun 2, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [ICLR'25 Oral] MMIE: Massive Multimodal Interleaved Comprehension Benchmark for Large Vision-Language Models☆35Nov 3, 2024Updated last year
- Code for CVPR 2024 Oral "Neural Lineage"☆17Jun 18, 2024Updated 2 years ago
- The Classification of 105 Celebrities with Face-Recognition using Tensorflow-Framework☆12Mar 24, 2021Updated 5 years ago
- CoCo: Code as CoT for Text-to-Image Preview and Rare Concept Generation☆56Aug 7, 2026Updated 3 weeks ago
- Breaking Certifiable Defenses☆17Nov 22, 2022Updated 3 years ago
- Official Code for GPIC: A Giant Permissive Image Corpus for Visual Generation☆54Jun 4, 2026Updated 2 months ago
- Learning from Indirect Observations☆11Jul 16, 2021Updated 5 years ago
- This is a repository for awesome any2any work collection.☆30Jul 10, 2026Updated last month
- Q-resafe:Assessing Safety Risks and Quantization-aware Safety Patching for Quantized Large Language Models (ICML'2025)☆16Jun 28, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- code for EMNLP 2024 paper: How do Large Language Models Learn In-Context? Query and Key Matrices of In-Context Heads are Two Towers for M…☆13Nov 17, 2024Updated last year
- [ICCV 2025] Official code for paper: Beyond Text-Visual Attention: Exploiting Visual Cues for Effective Token Pruning in VLMs☆84Jul 1, 2025Updated last year
- Repository for 2030 project☆15Dec 29, 2025Updated 8 months ago
- ☆16Dec 25, 2025Updated 8 months ago
- [ACL2025 Findings] Migician: Revealing the Magic of Free-Form Multi-Image Grounding in Multimodal Large Language Models☆91May 20, 2025Updated last year
- [CCS'22] SSLGuard: A Watermarking Scheme for Self-supervised Learning Pre-trained Encoders☆18Jul 12, 2022Updated 4 years ago
- ☆22Jan 22, 2026Updated 7 months ago
- [WACV 2024] Instruct Me More! Random Prompting for Visual In-Context Learning☆18May 7, 2025Updated last year
- Official implementation for the paper "Video-Based Reward Modeling for Computer-Use Agents"☆17Mar 14, 2026Updated 5 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [ACL2026 oral] Uni-MMMU : A Massive Multi-discipline Multimodal Unified Benchmark☆27Apr 13, 2026Updated 4 months ago
- A simple PyTorch implementation of Learning Instance Activation Maps for Weakly Supervised Instance Segmentation, in CVPR 2019☆11Jun 18, 2020Updated 6 years ago
- ComfyUI version of WithAnyone☆26Dec 18, 2025Updated 8 months ago
- ☆15Jul 30, 2020Updated 6 years ago
- Fine-tuning dino v2 for semantic segmentation task on MSCOCO.☆31Jun 6, 2023Updated 3 years ago
- codes for AAAI papar : Infrared-Visible Cross-Modal Person Re-Identification with an X Modality☆19Apr 8, 2020Updated 6 years ago
- The official implementation of the paper SAEdit: Token-level control for continuous image editing via Sparse AutoEncoder☆24Oct 19, 2025Updated 10 months ago
- MACER: MAximizing CErtified Radius (ICLR 2020)☆31Jan 5, 2020Updated 6 years ago
- Autonomous Knowledge Acquisition and Reasearch Intelligence☆47Mar 8, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆21Oct 1, 2024Updated last year
- [ICLR 2026] The official repository for paper "ThinkMorph: Emergent Properties in Multimodal Interleaved Chain-of-Thought Reasoning"☆193May 1, 2026Updated 3 months ago
- ☆24Nov 11, 2022Updated 3 years ago
- [CCS 2021] TSS: Transformation-specific smoothing for robustness certification☆26Oct 3, 2023Updated 2 years ago
- ☆32Jun 13, 2026Updated 2 months ago
- [CVPR'26] When Visualizing is the First Step to Reasoning: MIRA, a Benchmark for Visual Chain-of-Thought☆32Feb 14, 2026Updated 6 months ago
- [CVPR 2025] PACT: Pruning and Clustering-Based Token Reduction for Faster Visual Language Models☆60Jan 30, 2026Updated 7 months ago