MMSci: A Multimodal Multi-Discipline Dataset for PhD-Level Scientific Comprehension
☆51Dec 3, 2024Updated last year
Alternatives and similar repositories for MMSci
Users that are interested in MMSci are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆15Jul 24, 2024Updated 2 years ago
- Serializing molecule 3D structures☆14Nov 27, 2024Updated last year
- ☆11Jan 3, 2024Updated 2 years ago
- ☆15Dec 4, 2023Updated 2 years ago
- ChartSum is a large scale benchmark for automatic chart to text summarization☆11Jul 20, 2023Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆25May 16, 2024Updated 2 years ago
- 操作系统内存管理项目☆14Jun 5, 2021Updated 5 years ago
- This repository is maintained to release dataset and models for multimodal puzzle reasoning.☆117Feb 26, 2025Updated last year
- SciAssess is a comprehensive benchmark for evaluating Large Language Models' proficiency in scientific literature analysis across various…☆89May 21, 2025Updated last year
- Downscaling Intelligence: Exploring Perception and Reasoning Bottlenecks in Small Multimodal Models☆25Mar 21, 2026Updated 4 months ago
- This repo contains code and data for ICLR 2025 paper MIA-Bench: Towards Better Instruction Following Evaluation of Multimodal LLMs☆38Mar 9, 2025Updated last year
- Repository for the KVP10k dataset☆23Sep 18, 2025Updated 10 months ago
- [EMNLP 2024] Official code for "Beyond Embeddings: The Promise of Visual Table in Multi-Modal Models"☆20Oct 17, 2024Updated last year
- A Vision-Language Benchmark for Microscopy Understanding☆31Mar 13, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- COLA: Evaluate how well your vision-language model can Compose Objects Localized with Attributes!☆25May 14, 2026Updated 2 months ago
- ☆23Aug 27, 2025Updated 11 months ago
- Code release for "SPIQA: A Dataset for Multimodal Question Answering on Scientific Papers" [NeurIPS D&B, 2024]☆76Jan 13, 2025Updated last year
- ☆41Sep 9, 2025Updated 10 months ago
- Official implementation of the ECCV2024 paper: Generalizable Facial Expression Recognition☆23Sep 20, 2024Updated last year
- ☆15Jan 9, 2026Updated 6 months ago
- [ICLR 2025] MMFakeBench: A Mixed-Source Multimodal Misinformation Detection Benchmark for LVLMs☆55Mar 25, 2025Updated last year
- ☆88Aug 18, 2024Updated last year
- UniParser-Tools: SDKs, Utilities, and Post-Processing for Industrial-Grade Multi-Modal PDF Parsing☆22Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆42Jul 15, 2025Updated last year
- BioT5 (EMNLP 2023) and BioT5+ (ACL 2024 Findings)☆127Sep 14, 2024Updated last year
- Official implementation of EMNLP'2022 paper "Non-Parametric Domain Adaptation for End-to-End Speech Translation"☆11Oct 26, 2022Updated 3 years ago
- SciKnowEval: Evaluating Multi-level Scientific Knowledge of Large Language Models☆30Jul 13, 2025Updated last year
- ☆23Feb 3, 2026Updated 5 months ago
- ☆19Sep 11, 2024Updated last year
- ☆47Nov 8, 2024Updated last year
- ☆27Jul 3, 2024Updated 2 years ago
- ☆27Jun 11, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Source code of our EMNLP 2024 paper "FactAlign: Long-form Factuality Alignment of Large Language Models"☆19Oct 3, 2024Updated last year
- Code of AAAI2025 Paper 《VIoTGPT: Learning to Schedule Vision Tools in LLMs towards Intelligent Video Internet of Things》☆16Jan 16, 2025Updated last year
- Jupyter notebooks for analysis and figures related to the native organelle IP paper☆14Mar 10, 2026Updated 4 months ago
- The official implementation of LinkerNet: Fragment Poses and Linker Co-Design with 3D Equivariant Diffusion (NeurIPS 2023 Spotlight)☆19Feb 23, 2024Updated 2 years ago
- ☆41Dec 7, 2025Updated 7 months ago
- SafeSora is a human preference dataset designed to support safety alignment research in the text-to-video generation field, aiming to enh…☆35Aug 20, 2024Updated last year
- Official implementation of SIGIR 2022 Paper "Task-Oriented Dialogue System as Natural Language Generation".☆14Apr 6, 2022Updated 4 years ago