A consistent Med-VQA dataset, C-SLAKE , extended by Slake for further consistency assessment .
☆13Jan 12, 2024Updated 2 years ago
Alternatives and similar repositories for CSLAKE
Users that are interested in CSLAKE are comparing it to the libraries listed below
Sorting:
- Multigranularity Contrastive cross-modal collaborative Generation (MCG) model for Video QA☆11Dec 13, 2023Updated 2 years ago
- Consistency Conditioned Memory Augmented Dynamic Diagnosis Model for Medical Visual Question Answering☆13Jan 12, 2024Updated 2 years ago
- Adapter-Enhanced Hierarchical Cross-Modal Pre-training for Lightweight Medical Report Generation☆12Jan 25, 2025Updated last year
- Observation Driven Memory Synergistic Planning for Continuous Vision-Language Navigation☆28Jun 14, 2024Updated last year
- [NeurIPS D&B'24]Enhancing vision-language models for medical imaging: bridging the 3D gap with innovative slice selection☆19Nov 25, 2024Updated last year
- The official implementation of "Surface Depth Estimation from Multi-view Stereo Satellite Images with Distribution Contrast Network”☆10May 16, 2025Updated 9 months ago
- The official repository for "One Model to Rule them All: Towards Universal Segmentation for Medical Images with Text Prompts"☆10Aug 16, 2024Updated last year
- The repo of the paper: Generalist Vision Foundation Models for Medical Imaging: A Case Study of Segment Anything Model on Zero-Shot Medic…☆11May 26, 2023Updated 2 years ago
- PyTorch implementation of video captioning☆13Sep 24, 2017Updated 8 years ago
- ☆13Oct 15, 2025Updated 4 months ago
- The official implementation of "Feature Distribution Normalization Network for Multi-View Stereo”.☆13Mar 5, 2025Updated 11 months ago
- Code implementation of RP3D-Diag☆17Nov 25, 2024Updated last year
- Code for our EMNLP-2022 paper: "Towards Robust Visual Question Answering: Making the Most of Biased Samples via Contrastive Learning"☆16Feb 22, 2023Updated 3 years ago
- [IEEE TMI'22] VQAMix: Conditional Triplet Mixup for Medical Visual Question Answering☆16Oct 9, 2022Updated 3 years ago
- ☆18Oct 13, 2022Updated 3 years ago
- [Science Advances] Demographic Bias of Vision-Language Foundation Models in Medical Imaging☆21Mar 28, 2025Updated 11 months ago
- A Layered Memory Network for MovieQA☆16Apr 27, 2018Updated 7 years ago
- Medical Knowledge-Based Network For Patient-oriented Visual Question Answering☆18Feb 25, 2023Updated 3 years ago
- ☆17Jul 21, 2022Updated 3 years ago
- AIOZ AI - Overcoming Data Limitation in Medical Visual Question Answering (MICCAI 2019)☆69Oct 3, 2023Updated 2 years ago
- The code of IJCAI2022 paper, Declaration-based Prompt Tuning for Visual Question Answering☆20May 10, 2022Updated 3 years ago
- ☆18Aug 29, 2025Updated 6 months ago
- ☆20Nov 25, 2024Updated last year
- Tracking the latest and greatest research papers on text-to-image generation.☆52Dec 2, 2025Updated 2 months ago
- A framework for Longitudinal Radiology Report Generation☆26Aug 10, 2024Updated last year
- PyTorch code for ROLL, a knowledge-based video story question answering model.☆21Sep 29, 2020Updated 5 years ago
- Repository of paper Consistency-preserving Visual Question Answering in Medical Imaging (MICCAI2022)☆25Mar 28, 2023Updated 2 years ago
- Diagnostic Captioning☆25Dec 8, 2022Updated 3 years ago
- Repository of our accepted CVPR2022 paper "Counterfactual Cycle-Consistent Learning for Instruction Following and Generation in Vision-La…☆28Mar 4, 2022Updated 3 years ago
- The system detects players and the ball with YOLO, assigns teams via zero-shot jersey classification, tracks ball possession, maps court …☆36Jul 25, 2025Updated 7 months ago
- ☆36Jan 9, 2026Updated last month
- The first ophthalmology Large Language-and-Vision Assistant based on Instructions and Dialogue☆34Oct 1, 2024Updated last year
- Official Code for LightVLA (ICRA 2026)☆78Jan 31, 2026Updated last month
- ☆25Sep 8, 2017Updated 8 years ago
- ROCK model for Knowledge-Based VQA in Videos☆31Oct 19, 2020Updated 5 years ago
- code for Expert Knowledge-Aware Image Difference Graph Representation Learning for Difference-Aware Medical Visual Question Answering☆29May 30, 2025Updated 9 months ago
- Official repository for the paper "Rad-ReStruct: A Novel VQA Benchmark and Method for Structured Radiology Reporting" (MICCAI23)☆32Jan 4, 2024Updated 2 years ago
- Action Proposals generated by deep models☆29Mar 19, 2017Updated 8 years ago
- Traces the boundary of a set of points belonging to an aerial LiDAR scan of a building (part).☆34Jan 9, 2023Updated 3 years ago