[ACL 2025] Exploring Compositional Generalization of Multimodal LLMs for Medical Imaging
☆40Jun 4, 2025Updated last year
Alternatives and similar repositories for Med-MAT
Users that are interested in Med-MAT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR'25] ApolloMoE: Efficiently Democratizing Medical LLMs for 50 Languages via a Mixture of Language Family Experts☆53Nov 20, 2024Updated last year
- Agentifying Patient Dynamics within LLMs through Interacting with Clinical World Model☆31May 15, 2026Updated 2 months ago
- MedGen: Unlocking Medical Video Generation by Scaling Granularly-annotated Medical Videos.☆33Apr 18, 2026Updated 3 months ago
- The repository of the ACCV 2024 paper "FG-CXR: A Radiologist-Aligned Gaze Dataset for Enhancing Interpretability in Chest X-Ray Report Ge…☆12Jul 28, 2025Updated last year
- Incentivizing Multimodal Complex Reasoning in Dentistry☆38Jun 29, 2026Updated last month
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- This repository includes the full release of PrinciplismQA dataset and assessment scripts.☆16Jul 14, 2026Updated 3 weeks ago
- GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI.☆101Jun 23, 2026Updated last month
- Reproduction of the complete process of DeepSeek-R1 on small-scale models, including Pre-training, SFT, and RL.☆30Mar 11, 2025Updated last year
- The official codes for "Can Modern LLMs Act as Agent Cores in Radiology Environments?"☆29Jan 22, 2025Updated last year
- ☆24Jan 11, 2025Updated last year
- ☆18Jul 13, 2026Updated 3 weeks ago
- [ 🎯 NeurIPS 2025 ] 3D-RAD 🩻: A Comprehensive 3D Radiology Med-VQA Dataset with Multi-Temporal Analysis and Diverse Diagnostic Tasks☆34Jun 22, 2026Updated last month
- ☆25Nov 27, 2025Updated 8 months ago
- Towards Fine-grained Audio Captioning with Multimodal Contextual Cues☆88Jan 4, 2026Updated 7 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆49Jul 17, 2026Updated 3 weeks ago
- Fast LLM Training CodeBase With dynamic strategy choosing [Deepspeed+Megatron+FlashAttention+CudaFusionKernel+Compiler];☆41Jan 4, 2024Updated 2 years ago
- [NAACL 2025] VividMed: Vision Language Model with Versatile Visual Grounding for Medicine☆32Mar 10, 2025Updated last year
- The official repository of the paper 'Towards a Multimodal Large Language Model with Pixel-Level Insight for Biomedicine'☆134Jul 7, 2026Updated last month
- Encourage Medical LLM to engage in deep thinking similar to DeepSeek-R1.☆26Apr 24, 2025Updated last year
- The official implementation of "Enhancing Representation in Radiography-Reports Foundation Model: A Granular Alignment Algorithm Using Ma…☆12Sep 13, 2024Updated last year
- The "GPT-API-Accelerate" project provides a set of Python classes for accelerating the process of generating responses to prompts using t…☆23Oct 12, 2024Updated last year
- A Python tool to evaluate the performance of VLM on the medical domain.☆89Aug 5, 2025Updated last year
- RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography☆45Jun 3, 2026Updated 2 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [ISBI 2025] Design Data Before Models: Using large vision-language models to automatically enhance medical dataset annotations.☆35Jan 28, 2026Updated 6 months ago
- ☆30Oct 13, 2025Updated 9 months ago
- GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI.☆87Dec 17, 2024Updated last year
- ☆20Jan 3, 2025Updated last year
- CVPR2026☆36Sep 18, 2025Updated 10 months ago
- ☆140Nov 13, 2025Updated 8 months ago
- Official implementation of "MedITok: A Unified Tokenizer for Medical Image Synthesis and Interpretation"☆30Apr 3, 2026Updated 4 months ago
- ECCV[2024] "Modelling Competitive Behaviors in Autonomous Driving Under Generative World Model" official implement☆17Jul 15, 2025Updated last year
- ☆73Feb 3, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This repository contains codes for AortaSeg24 Grand-Challenge.☆21Nov 22, 2024Updated last year
- NeurIPS[2023] "Multi-Modal Inverse Constrained Reinforcement Learning from a Mixture of Demonstrations" official implement☆13Feb 19, 2024Updated 2 years ago
- [EMNLP 2024] This is the code for our paper "BMRetriever: Tuning Large Language Models as Better Biomedical Text Retrievers".☆26Sep 19, 2024Updated last year
- LLaVa Version of RaDialog☆26May 27, 2025Updated last year
- [MICCAI 2026] A longitudinal, multimodal algorithm for multi-tumor segmentation (learning from reports).☆15Jun 29, 2026Updated last month
- An interpretable large language model (LLM) for medical diagnosis.☆164Sep 12, 2024Updated last year
- IEEE JBHI 2026 | MedSegAgent: A Universal and Scalable Multi-Agent System for Instructive Medical Image Segmentation☆46May 20, 2026Updated 2 months ago