☆48Sep 5, 2024Updated last year
Alternatives and similar repositories for CMMMU
Users that are interested in CMMMU are comparing it to the libraries listed below
Sorting:
- Accelerating the development of large multimodal models (LMMs) with lmms-eval☆14Oct 14, 2024Updated last year
- ☆54Mar 19, 2025Updated last year
- This repo contains evaluation code for the paper "MMMU: A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for E…☆549Feb 12, 2026Updated last month
- [ECCV 2024] "REVISION: Rendering Tools Enable Spatial Fidelity in Vision-Language Models"☆13Aug 6, 2024Updated last year
- ☆19Nov 12, 2024Updated last year
- Official repo for StableLLAVA☆95Dec 22, 2023Updated 2 years ago
- ☆54Nov 14, 2024Updated last year
- ☆39Aug 9, 2022Updated 3 years ago
- This repo contains the code for "MEGA-Bench Scaling Multimodal Evaluation to over 500 Real-World Tasks" [ICLR 2025]☆79Jul 1, 2025Updated 8 months ago
- ☆19Aug 3, 2024Updated last year
- 学术主页 | Academic Page☆14Updated this week
- [CVPR 2024] DiffAgent: Fast and Accurate Text-to-Image API Selection with Large Language Model☆19Apr 16, 2024Updated last year
- Resources for paper "DialSummEval: Revisiting summarization evaluation for dialogues"☆15Jul 22, 2025Updated 7 months ago
- [ICCV2023] NoiseDet: Learning from Noisy Data for Semi-Superivsed 3D Object Detection☆20Feb 5, 2023Updated 3 years ago
- ☆32Feb 8, 2024Updated 2 years ago
- ☆125Jul 29, 2024Updated last year
- ☆19Dec 6, 2023Updated 2 years ago
- A method for evaluating the high-level coherence of machine-generated texts. Identifies high-level coherence issues in transformer-based …☆11Mar 18, 2023Updated 3 years ago
- MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities (ICML 2024)☆323Jan 20, 2025Updated last year
- ☆36Sep 6, 2024Updated last year
- Official github repo of G-LLaVA☆148Feb 20, 2025Updated last year
- Official Repo for the paper: VCR: Visual Caption Restoration. Check arxiv.org/pdf/2406.06462 for details.☆32Feb 26, 2025Updated last year
- [ICLR2026] VideoChat-Flash: Hierarchical Compression for Long-Context Video Modeling☆511Nov 18, 2025Updated 4 months ago
- 一个mmcv 的logger hook, 可以用来把模型结果推送到微信上☆21Oct 11, 2022Updated 3 years ago
- Official code for the paper, "TaCA: Upgrading Your Visual Foundation Model with Task-agnostic Compatible Adapter".☆16Jun 20, 2023Updated 2 years ago
- MTVQA: Benchmarking Multilingual Text-Centric Visual Question Answering. A comprehensive evaluation of multimodal large model multilingua…☆64May 15, 2025Updated 10 months ago
- Implementation for "The Scalability of Simplicity: Empirical Analysis of Vision-Language Learning with a Single Transformer"☆80Oct 29, 2025Updated 4 months ago
- ☆27Jan 23, 2024Updated 2 years ago
- Extending Conformal Prediction to LLMs☆69Jun 21, 2024Updated last year
- [ICLR'24] Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning☆296Mar 13, 2024Updated 2 years ago
- [ICML 2024] | MMT-Bench: A Comprehensive Multimodal Benchmark for Evaluating Large Vision-Language Models Towards Multitask AGI☆117Jul 18, 2024Updated last year
- ☆28Feb 10, 2025Updated last year
- [ICLR 2024 & ECCV 2024] The All-Seeing Projects: Towards Panoptic Visual Recognition&Understanding and General Relation Comprehension of …☆506Aug 9, 2024Updated last year
- Official repo for "SoftMatch: Addressing the Quantity-Quality Trade-off in Semi-Supervised Learning", accepted by ICLR 2023.☆22Jan 31, 2023Updated 3 years ago
- [ICML 2025] This is the official PyTorch implementation of "OmniBal: Towards Fast Instruction-Tuning for Vision-Language Models via Omniv…☆27Jun 16, 2025Updated 9 months ago
- ☆27Mar 21, 2024Updated last year
- This is for C2D2 Dataset: A Resource for Analyzing Cognitive Distortions and Its Impact on Mental Health☆33Nov 10, 2023Updated 2 years ago
- ☆17Feb 22, 2024Updated 2 years ago
- [CVPR 2024] Data and benchmark code for the EgoExoLearn dataset☆81Aug 26, 2025Updated 6 months ago