A consistent Med-VQA dataset, C-SLAKE , extended by Slake for further consistency assessment .
☆17Jan 12, 2024Updated 2 years ago
Alternatives and similar repositories for CSLAKE
Users that are interested in CSLAKE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Consistency Conditioned Memory Augmented Dynamic Diagnosis Model for Medical Visual Question Answering☆16Jan 12, 2024Updated 2 years ago
- Multigranularity Contrastive cross-modal collaborative Generation (MCG) model for Video QA☆12Dec 13, 2023Updated 2 years ago
- Adapter-Enhanced Hierarchical Cross-Modal Pre-training for Lightweight Medical Report Generation☆15Jan 25, 2025Updated last year
- VisionDreamer: High-Fidelity Text-to-3D Generation via Mesh-Guided 3D Gaussian Splatting☆19Jul 7, 2025Updated last year
- ☆16Sep 17, 2025Updated 10 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [IEEE JSTARS] The official implementation of "Surface Depth Estimation from Multi-view Stereo Satellite Images with Distribution Contrast…☆13May 16, 2025Updated last year
- ☆13Oct 15, 2025Updated 9 months ago
- The official repository for "One Model to Rule them All: Towards Universal Segmentation for Medical Images with Text Prompts"☆10Aug 16, 2024Updated last year
- ☆10Oct 20, 2022Updated 3 years ago
- ☆17Jul 21, 2022Updated 4 years ago
- [IEEE TMI'22] VQAMix: Conditional Triplet Mixup for Medical Visual Question Answering☆16Oct 9, 2022Updated 3 years ago
- Code implementation of RP3D-Diag☆17Nov 25, 2024Updated last year
- Code for our EMNLP-2022 paper: "Towards Robust Visual Question Answering: Making the Most of Biased Samples via Contrastive Learning"☆16Feb 22, 2023Updated 3 years ago
- ☆19Oct 13, 2022Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Code and data for paper "Exploring Hallucination of Large Multimodal Models in Video Understanding: Benchmark, Analysis and Mitigation".☆25Oct 22, 2025Updated 9 months ago
- Repository of our accepted CVPR2022 paper "Counterfactual Cycle-Consistent Learning for Instruction Following and Generation in Vision-La…☆28Mar 4, 2022Updated 4 years ago
- PyTorch implementation of video captioning☆13Sep 24, 2017Updated 8 years ago
- Medical Knowledge-Based Network For Patient-oriented Visual Question Answering☆19Feb 25, 2023Updated 3 years ago
- ☆24Apr 12, 2026Updated 3 months ago
- Repository of paper Consistency-preserving Visual Question Answering in Medical Imaging (MICCAI2022)☆26Mar 28, 2023Updated 3 years ago
- AIOZ AI - Overcoming Data Limitation in Medical Visual Question Answering (MICCAI 2019)☆70Apr 21, 2026Updated 3 months ago
- [ICCV 2025] Official repository of "Mitigating Object Hallucinations via Sentence-Level Early Intervention".☆31Jul 2, 2026Updated last month
- A Layered Memory Network for MovieQA☆16Apr 27, 2018Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆39Mar 19, 2026Updated 4 months ago
- PyTorch code for ROLL, a knowledge-based video story question answering model.☆21Sep 29, 2020Updated 5 years ago
- Code for CVPR 2023 paper "Procedure-Aware Pretraining for Instructional Video Understanding"☆50Jun 2, 2026Updated 2 months ago
- Traces the boundary of a set of points belonging to an aerial LiDAR scan of a building (part).☆34Jan 9, 2023Updated 3 years ago
- ☆44Oct 20, 2023Updated 2 years ago
- Deep Multimodal Neural Architecture Search☆29Nov 15, 2020Updated 5 years ago
- ☆20Nov 25, 2024Updated last year
- The first ophthalmology Large Language-and-Vision Assistant based on Instructions and Dialogue☆39Oct 1, 2024Updated last year
- [ESWA'26] Factual Serialization Enhancement: A Key Innovation for Chest X-ray Report Generation☆19Apr 25, 2026Updated 3 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- ☆19Aug 29, 2025Updated 11 months ago
- ROCK model for Knowledge-Based VQA in Videos☆31Oct 19, 2020Updated 5 years ago
- Diagnostic Captioning☆25Dec 8, 2022Updated 3 years ago
- This repository is made for the paper: Masked Vision and Language Pre-training with Unimodal and Multimodal Contrastive Losses for Medica…☆50Jul 10, 2024Updated 2 years ago
- Action Proposals generated by deep models☆29Mar 19, 2017Updated 9 years ago
- ☆50Apr 27, 2024Updated 2 years ago
- [CVPR 2022] X-Trans2Cap: Cross-Modal Knowledge Transfer using Transformer for 3D Dense Captioning☆36Aug 26, 2022Updated 3 years ago