MixGen: A New Multi-Modal Data Augmentation
☆126Jan 9, 2023Updated 3 years ago
Alternatives and similar repositories for mix-generation
Users that are interested in mix-generation are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- PyTorch code for "VL-Adapter: Parameter-Efficient Transfer Learning for Vision-and-Language Tasks" (CVPR2022)☆212Dec 18, 2022Updated 3 years ago
- A library of transformer models for computer vision and multi-modality research☆49Sep 7, 2021Updated 4 years ago
- Official code for the paper, "TaCA: Upgrading Your Visual Foundation Model with Task-agnostic Compatible Adapter".☆16Jun 20, 2023Updated 3 years ago
- [ECCV'22 Poster] Explicit Image Caption Editing☆22Nov 30, 2022Updated 3 years ago
- TextAdaIN: Paying Attention to Shortcut Learning in Text Recognizers☆21Jul 26, 2022Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- code for TCL: Vision-Language Pre-Training with Triple Contrastive Learning, CVPR 2022☆270Oct 2, 2024Updated last year
- Optimized code based on M2 for faster image captioning training☆21Nov 18, 2022Updated 3 years ago
- The imdb files with SBD-Trans OCR for TextVQA dataset.☆11Nov 30, 2021Updated 4 years ago
- [CVPR 2023] Official repository of paper titled "MaPLe: Multi-modal Prompt Learning".☆818Jul 24, 2023Updated 3 years ago
- Code for ALBEF: a new vision-language pre-training method☆1,755Sep 20, 2022Updated 3 years ago
- Code Example for Learning Multimodal Data Augmentation in Feature Space☆44Mar 11, 2023Updated 3 years ago
- [CVPR 2022 Oral] Crafting Better Contrastive Views for Siamese Representation Learning☆289Jun 27, 2022Updated 4 years ago
- SVL-Adapter: Self-Supervised Adapter for Vision-Language Pretrained Models☆21Jan 11, 2024Updated 2 years ago
- Official Code of ECCV 2022 paper MS-CLIP☆90Jul 27, 2022Updated 4 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A subset of YFCC100M. Tools, checking scripts and links of web drive to download datasets(uncompressed).☆19Aug 5, 2026Updated 3 weeks ago
- Tensorflow implementation of "Hide-and-Seek: Forcing a Network to be Meticulous for Weakly-supervised Object and Action Localization"[ICC…☆13Mar 29, 2019Updated 7 years ago
- Official code for the paper "Contrast and Classify: Training Robust VQA Models" published at ICCV, 2021☆19Jul 27, 2021Updated 5 years ago
- Learning Debiased and Disentangled Representations for Semantic Segmentation (NeurIPS 2021)☆13Jan 23, 2022Updated 4 years ago
- BigDetection: A Large-scale Benchmark for Improved Object Detector Pre-training☆399Oct 23, 2024Updated last year
- Official implementation for the paper "Prompt Pre-Training with Over Twenty-Thousand Classes for Open-Vocabulary Visual Recognition"☆259May 3, 2024Updated 2 years ago
- PyTorch code for MUST☆108May 1, 2025Updated last year
- FreeVA: Offline MLLM as Training-Free Video Assistant☆69Jun 9, 2024Updated 2 years ago
- Official implementation for "Parameter-Efficient Fine-Tuning Design Spaces"☆27Jan 4, 2023Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆11Feb 9, 2026Updated 6 months ago
- ☆37May 7, 2023Updated 3 years ago
- [ICLR2024] The official implementation of paper "UniAdapter: Unified Parameter-Efficient Transfer Learning for Cross-modal Modeling", by …☆77Jan 27, 2024Updated 2 years ago
- GRIT: Faster and Better Image-captioning Transformer (ECCV 2022)☆199May 9, 2023Updated 3 years ago
- Towards Local Visual Modeling for Image Captioning☆30Mar 31, 2023Updated 3 years ago
- [CVPR2023] The code for 《Position-guided Text Prompt for Vision-Language Pre-training》☆149Jun 7, 2023Updated 3 years ago
- Incremental Generative Occlusion Adversarial Suppression Network for Person ReID (IEEE T-IP 2021)☆14Dec 1, 2023Updated 2 years ago
- ☆14Jul 21, 2022Updated 4 years ago
- ☆23Mar 19, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Supervision Exists Everywhere: A Data Efficient Contrastive Language-Image Pre-training Paradigm☆678Sep 19, 2022Updated 3 years ago
- ☆24Apr 4, 2022Updated 4 years ago
- Align and Prompt: Video-and-Language Pre-training with Entity Prompts☆188May 1, 2025Updated last year
- Video + CLIP Baseline for Ego4D Long Term Action Anticipation Challenge (CVPR 2022)☆15Jul 4, 2022Updated 4 years ago
- ☆22Nov 23, 2023Updated 2 years ago
- [ECCV2022] Contrastive Vision-Language Pre-training with Limited Resources☆46Sep 29, 2022Updated 3 years ago
- PyTorch codes for "LST: Ladder Side-Tuning for Parameter and Memory Efficient Transfer Learning"☆241Jan 20, 2023Updated 3 years ago