SimVLM ---SIMPLE VISUAL LANGUAGE MODEL PRETRAINING WITH WEAK SUPERVISION
☆36Nov 7, 2022Updated 3 years ago
Alternatives and similar repositories for SimVLM
Users that are interested in SimVLM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for GHA (ACCV2018)☆13Oct 31, 2018Updated 7 years ago
- A PyTorch implementation of Multimodal Few-Shot Learning with Frozen Language Models with OPT.☆43Jul 23, 2022Updated 4 years ago
- ☆12Oct 12, 2024Updated last year
- Code for the ACL 2023 paper Scene Graph as Pivoting: Inference-time Image-free Unsupervised Multimodal Machine Translation with Visual Sc…☆12May 19, 2023Updated 3 years ago
- A NVIDIA GPU monitor web tool☆10Jul 6, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Domain Adaptation as a Problem of Inference on Graphical Models☆29Dec 23, 2020Updated 5 years ago
- This repository is the official implementation of the aaai2022 paper "Zero-Shot Out-of-Distribution Detection Based on the Pre-trained Mo…☆21Sep 6, 2023Updated 2 years ago
- Pytorch Implementation of CLIP-Lite | Accepted at AISTATS 2023☆14Mar 17, 2023Updated 3 years ago
- Implementation of the deepmind Flamingo vision-language model, based on Hugging Face language models and ready for training☆171Apr 27, 2023Updated 3 years ago
- Plug-and-Play Document Modules for Pre-trained Models☆25May 28, 2023Updated 3 years ago
- C4RepSet: Representative Subset from C4 data for Training Pre-trained LMs☆11Jan 13, 2023Updated 3 years ago
- Code for Learned Thresholds Token Merging and Pruning for Vision Transformers (LTMP). A technique to reduce the size of Vision Transforme…☆17Nov 24, 2024Updated last year
- GitHub Profile README☆10Updated this week
- Code for "On Long-Tailed Phenomena in NMT".☆10Jan 10, 2021Updated 5 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [Paper][ISWC 2021] Zero-shot Visual Question Answering using Knowledge Graph☆72Feb 9, 2024Updated 2 years ago
- ☆26Aug 14, 2022Updated 4 years ago
- Using pretrained encoder and language models to generate captions from multimedia inputs.☆101Mar 11, 2023Updated 3 years ago
- ☆20May 5, 2023Updated 3 years ago
- ☆12May 13, 2023Updated 3 years ago
- Dataset, metrics, and models for TACL 2023 paper MACSUM: Controllable Summarization with Mixed Attributes.☆34Jul 25, 2023Updated 3 years ago
- Official Implementation of LADS (Latent Augmentation using Domain descriptionS)☆50Apr 18, 2023Updated 3 years ago
- The code repository for EMNLP 2021 paper "Vision Guided Generative Pre-trained Language Models for Multimodal Abstractive Summarization".☆57Jan 14, 2022Updated 4 years ago
- Source Code for TrustCom2022 Accepted Paper " 'Comments Matter and The More The Better': Improving Rumor Detecion with User Comments".☆19May 23, 2023Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆33Apr 23, 2023Updated 3 years ago
- Challenge the customer to shop task with tripletNet☆16Mar 30, 2020Updated 6 years ago
- Code release for Unsupervised Domain Adaptation via Distilled Discriminative Clustering published by Pattern Recognition in 2022☆11May 19, 2023Updated 3 years ago
- A multimodal dataset for google map restaurants.☆12Sep 28, 2022Updated 3 years ago
- aigc evals☆10Dec 2, 2023Updated 2 years ago
- OptimSeed - Seed Word Selection for Weakly-Supervised Text Classification [NAACL SRW 2021]☆14Mar 29, 2021Updated 5 years ago
- Reproducing the fashion image categorization and retrieval baseline approach from https://github.com/switchablenorms/DeepFashion2 / https…☆19Nov 2, 2021Updated 4 years ago
- ☆13Mar 14, 2025Updated last year
- [HCLT 2022] Korean sentence text similarity dataset using naver shopping review☆25Oct 20, 2022Updated 3 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆11Oct 5, 2020Updated 5 years ago
- 🏆 The 1st Place Solution for AICity2022 Challenge Track2: Natural Language-Based Vehicle Retrieval.☆12Jul 25, 2022Updated 4 years ago
- [CVPR'24] Official implementation of our paper "Self-Supervised Facial Representation Learning with Facial Region Awareness"☆15Mar 8, 2024Updated 2 years ago
- Coarse-to-Fine Reasoning for Visual Question Answering (CVPRW'22)☆48Apr 22, 2026Updated 4 months ago
- ☆12Sep 26, 2019Updated 6 years ago
- EDUVSUM is a multimodal neural architecture that utilizes state-of-the-art audio, visual and textual features to identify important tempo…☆23Mar 8, 2024Updated 2 years ago
- Microsoft question-answering dataset☆10Jun 16, 2023Updated 3 years ago