☆18Sep 13, 2023Updated 3 years ago
Alternatives and similar repositories for CIIC
Users that are interested in CIIC are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆18Nov 11, 2022Updated 3 years ago
- ☆18Oct 3, 2023Updated 3 years ago
- ☆30May 7, 2021Updated 5 years ago
- Code for Beyond Generic: Enhancing Image Captioning with Real-World Knowledge using Vision-Language Pre-Training Model☆13Feb 15, 2024Updated 2 years ago
- The code of paper "MLIP: Enhancing Medical Visual Representation with Divergence Encoder and Knowledge-guided Contrastive Learning" accep…☆10Mar 5, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ACM MM 2024] See or Guess: Counterfactually Regularized Image Captioning☆16Feb 17, 2025Updated last year
- Improving Medical Vision-Language Contrastive Pretraining with Semantics-aware Triage☆12Jun 25, 2023Updated 3 years ago
- official implementation of "Med-Unic: unifying cross-lingual medical vision-language pre-training by diminishing bias"☆18Sep 22, 2023Updated 3 years ago
- Official Code for 'RSTNet: Captioning with Adaptive Attention on Visual and Non-Visual Words' (CVPR 2021)☆123Dec 17, 2022Updated 3 years ago
- ☆45Jul 31, 2021Updated 5 years ago
- Official Implementation of "CLEFT: Language-Image Contrastive Learning with Efficient Large Language Model and Prompt Fine-Tuning" on MIC…☆18Feb 12, 2025Updated last year
- Joint Embedding of Deep Visual and Semantic Features for Medical Image Report Generation☆19Nov 13, 2025Updated 10 months ago
- Progressive Transformer-Based Generation of Radiology Reports☆25Jan 5, 2025Updated last year
- Repository for Vision-and-Language Navigation via Causal Learning (Accepted by CVPR 2024)☆103Jun 4, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ICCV 2023] This is the Pytorch code for our paper "Self-Supervised Cross-View Representation Reconstruction for Change Captioning".☆20Sep 25, 2025Updated last year
- Video-CoM: Interactive Video Reasoning via Chain of Manipulations☆23Sep 5, 2026Updated last month
- ☆22Aug 1, 2023Updated 3 years ago
- Repository for an end-to-end image captioning method PTSN(ACM MM22).☆60Dec 11, 2022Updated 3 years ago
- Implementation of paper "Improving Image Captioning with Better Use of Caption"☆33Sep 15, 2020Updated 6 years ago
- Code & data accompanying the paper ["Unveiling Implicit Deceptive Patterns in Multi-modal Fake News via Neuro-Symbolic Reasoning"].☆13Dec 21, 2023Updated 2 years ago
- ☆79Oct 8, 2022Updated 4 years ago
- Visual Semantic Relatedness Dataset for Captioning. CVPRW 2023☆10Feb 27, 2024Updated 2 years ago
- Offical code of Unlocking the Power of Spatial and Temporal Information in Medical Multimodal Pre-training[ICML 2024]☆26May 31, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- MICCAI 22 accepted paper “TranSQ: Transformer-based Semantic Query for Medical Report Generation“ for medical report generation☆27Sep 3, 2025Updated last year
- ☆35Nov 22, 2022Updated 3 years ago
- [MICCAI 2021 (Oral)] Official code repository for "Variational Topic Inference for Chest X-Ray Report Generation"☆21Mar 7, 2022Updated 4 years ago
- Entity-Aware Dual Co-Attention Network for Fake News Detection, EACL 2023 Findings☆10Jun 11, 2023Updated 3 years ago
- NICE challenge 2023 Track2 2nd result(total 4th) (CVPR 2023) sponsered by LG AI/Shutterstock/SNU☆11Jun 22, 2023Updated 3 years ago
- [CVPR24 Highlights] Polos: Multimodal Metric Learning from Human Feedback for Image Captioning☆33Jun 12, 2026Updated 3 months ago
- EfficientVLM: Fast and Accurate Vision-Language Models via Knowledge Distillation and Modal-adaptive Pruning (ACL 2023)☆33Jul 18, 2023Updated 3 years ago
- [EACL'23] COVID-VTS: Fact Extraction and Verification on Short Video Platforms☆12Sep 26, 2023Updated 3 years ago
- [IJCAI 2023] CLE-ViT: Contrastive Learning Encoded Transformer for Ultra-Fine-Grained Visual Categorization.☆10Nov 3, 2023Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [ICCV-2023] Towards Unifying Medical Vision-and-Language Pre-training via Soft Prompts☆78Mar 22, 2024Updated 2 years ago
- [NeurIPS 2025] Reasoning MLLM, Share-GRPO, advantage vanishing, sparse reward☆38Sep 19, 2025Updated last year
- [CVPR2024] PairAug: What Can Augmented Image-Text Pairs Do for Radiology?☆29Nov 11, 2024Updated last year
- Code Repository for CausalDiffAE (ECAI 2024)☆26Oct 19, 2024Updated last year
- A Good Prompt Is Worth Millions of Parameters: Low-resource Prompt-based Learning for Vision-Language Models (ACL 2022)☆42May 13, 2022Updated 4 years ago
- Multi-Aspect Vision Language Pretraining - CVPR2024☆92Aug 20, 2024Updated 2 years ago
- An official implementation of Advancing Radiograph Representation Learning with Masked Record Modeling (ICLR'23)☆77Feb 21, 2023Updated 3 years ago