☆66Dec 15, 2023Updated 2 years ago
Alternatives and similar repositories for image-caption-baseline
Users that are interested in image-caption-baseline are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆60Nov 17, 2022Updated 3 years ago
- Bling's Object detection tool☆55Jan 9, 2023Updated 3 years ago
- Bridging Vision and Language Model☆287Mar 27, 2023Updated 3 years ago
- Latent dirichlet allocation using Sklearn☆18Aug 6, 2018Updated 8 years ago
- ☆17Oct 15, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Official code repo for "ProTo: program-guided Transformers for Program-guided Tasks☆21Apr 15, 2022Updated 4 years ago
- K-PLUG: Knowledge-injected Pre-trained Language Model for Natural Language Understanding and Generation in E-Commerce (Findings of EMNLP …☆30Jan 6, 2023Updated 3 years ago
- Implementation of the Benchmark Approaches for Medical Instructional Video Classification (MedVidCL) and Medical Video Question Answering…☆31Jan 31, 2023Updated 3 years ago
- Train a model for Image Caption from ViT and GPT pretrained model☆18Mar 25, 2023Updated 3 years ago
- Cross-lingual image captioning☆93May 9, 2022Updated 4 years ago
- Testing prompts with SDXL☆16Jul 28, 2023Updated 3 years ago
- WuDaoMM this is a data project☆75Apr 29, 2022Updated 4 years ago
- Enriching MS-COCO with Chinese sentences and tags for cross-lingual multimedia tasks☆215Feb 12, 2025Updated last year
- 格物-多语言和中文大规模预训练模型-轻量版,涵盖纯中文、知识增强、113个语种多语言,采用主流Roberta架构,适用于NLU和NLG任务, 支持pytorch、tensorflow、uer、huggingface等框架。 Multilingual and Chinese …☆30Nov 17, 2022Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- kdexd/coco-caption@de6f385☆26Apr 21, 2020Updated 6 years ago
- The dataset for paper "Why Do We Click: Visual Impression-aware News Recommendation", ACM MM 2021☆16Feb 24, 2022Updated 4 years ago
- A subset of YFCC100M. Tools, checking scripts and links of web drive to download datasets(uncompressed).☆19Aug 5, 2026Updated last month
- ROSITA: Enhancing Vision-and-Language Semantic Alignments via Cross- and Intra-modal Knowledge Integration☆57Jun 13, 2023Updated 3 years ago
- Code for ALBEF: a new vision-language pre-training method☆1,754Sep 20, 2022Updated 4 years ago
- Visual Semantic Relatedness Dataset for Captioning. CVPRW 2023☆10Feb 27, 2024Updated 2 years ago
- ☆40Nov 23, 2022Updated 3 years ago
- Code and dataset for our Bioinformatics 2022 paper: "A Benchmark for Automatic Medical Consultation System: Frameworks, Tasks and Datase…☆69Dec 24, 2022Updated 3 years ago
- An interactive storybook built with the help of ChatGPT and Stable Diffusion.☆13Jun 28, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Using Low-rank adaptation to quickly fine-tune diffusion models.☆11Mar 14, 2023Updated 3 years ago
- ☆13Sep 15, 2022Updated 4 years ago
- ☆11Apr 19, 2021Updated 5 years ago
- Medical ML Benchmark☆11May 16, 2023Updated 3 years ago
- CVPR 2021 Official Pytorch Code for UC2: Universal Cross-lingual Cross-modal Vision-and-Language Pre-training☆34Nov 9, 2021Updated 4 years ago
- Data and models for Misinfo Reaction Frames paper.☆14Jun 9, 2024Updated 2 years ago
- ☆13May 23, 2025Updated last year
- Entity-Aware Dual Co-Attention Network for Fake News Detection, EACL 2023 Findings☆10Jun 11, 2023Updated 3 years ago
- ☆10Jun 1, 2018Updated 8 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Code for our ACL2021 paper: "Check It Again: Progressive Visual Question Answering via Visual Entailment"☆31Nov 24, 2021Updated 4 years ago
- Multi Task Vision and Language☆822Feb 16, 2022Updated 4 years ago
- [ICLR 2022] code for "How Much Can CLIP Benefit Vision-and-Language Tasks?" https://arxiv.org/abs/2107.06383☆420Oct 28, 2022Updated 3 years ago
- Math24o: 高中奥林匹克数学竞赛测评集 High School Olympiad Mathematics Chinese Benchmark☆14Mar 27, 2025Updated last year
- ☆11Mar 13, 2023Updated 3 years ago
- [EACL'23] COVID-VTS: Fact Extraction and Verification on Short Video Platforms☆12Sep 26, 2023Updated 3 years ago
- ☆170Nov 9, 2023Updated 2 years ago