☆25Jul 10, 2023Updated 3 years ago
Alternatives and similar repositories for semantic-image-text-alignment
Users that are interested in semantic-image-text-alignment are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- SotA text-only image/video method (IJCAI 2023)☆14Jan 9, 2024Updated 2 years ago
- [NeurIPS 2022] code for "K-LITE: Learning Transferable Visual Models with External Knowledge" https://arxiv.org/abs/2204.09222☆54Jun 12, 2023Updated 3 years ago
- Implementation of the "Learn No to Say Yes Better" paper.☆40Apr 3, 2026Updated 5 months ago
- Paper list of compositional zero-shot learning☆11Jul 5, 2022Updated 4 years ago
- ☆11Sep 7, 2020Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆13Jul 16, 2024Updated 2 years ago
- Code and data release for the paper "Learning Fine-grained View-Invariant Representations from Unpaired Ego-Exo Videos via Temporal Align…☆19Apr 5, 2024Updated 2 years ago
- [EMNLP 2021] Code and data for our paper "Vision-and-Language or Vision-for-Language? On Cross-Modal Influence in Multimodal Transformers…☆20Jan 17, 2022Updated 4 years ago
- PyTorch code for the Findings of EMNLP 2021 paper "Does Vision-and-Language Pretraining Improve Lexical Grounding?"☆11Sep 26, 2021Updated 4 years ago
- ☆14Jul 30, 2022Updated 4 years ago
- ☆34May 15, 2024Updated 2 years ago
- Publication sources, algorithm, code, result, conference poster, scientific paper for ICDAR, CIFED, VISAPP☆15Jul 8, 2022Updated 4 years ago
- Code and data release for the paper "Learning Object State Changes in Videos: An Open-World Perspective" (CVPR 2024)☆37Sep 9, 2024Updated last year
- some object detection algo☆14Jul 25, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆59Aug 7, 2023Updated 3 years ago
- ☆14Aug 19, 2024Updated 2 years ago
- Implementation of BIMRL: Brain Inspired Meta Reinforcement Learning - Roozbeh Razavi et al. (IROS 2022)☆10Dec 1, 2022Updated 3 years ago
- ☆13Apr 12, 2026Updated 4 months ago
- Code for paper: "Privately generating tabular data using language models".☆16Jun 13, 2023Updated 3 years ago
- Stable diffusion multi-model image matrix generator based on 🧨diffusers☆11Aug 11, 2024Updated 2 years ago
- Open-source strong baseline for domain generlization re-ID. We will udpate the strong baseline and CFD method~☆10Nov 30, 2021Updated 4 years ago
- Implementation of Retrieval-Augmented Denoising Diffusion Probabilistic Models in Pytorch☆66May 5, 2022Updated 4 years ago
- Official implementation and dataset for the NAACL 2024 paper "ComCLIP: Training-Free Compositional Image and Text Matching"☆37Aug 18, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- This repo is forked from https://github.com/deepinsight/insightface. I use their face detection (retinaface) and face recognition (arcfac…☆10Jan 26, 2024Updated 2 years ago
- Official implementation of "In-style: Bridging Text and Uncurated Videos with Style Transfer for Cross-modal Retrieval." ICCV 2023☆11Oct 5, 2023Updated 2 years ago
- Colorful Prompt Tuning for Pre-trained Vision-Language Models☆49Nov 1, 2022Updated 3 years ago
- CLiC: Concept Learning in Context☆10Jan 24, 2025Updated last year
- [ICLR 2024] Official repository for "Vision-by-Language for Training-Free Compositional Image Retrieval"☆89Jul 4, 2024Updated 2 years ago
- Code for the paper "AMEGO: Active Memory from long EGOcentric videos" published at ECCV 2024☆45Dec 7, 2024Updated last year
- [CVPR'21 Oral] Seeing Out of tHe bOx: End-to-End Pre-training for Vision-Language Representation Learning☆208Sep 30, 2022Updated 3 years ago
- ICCV23 "Householder Projector for Unsupervised Latent Semantics Discovery"☆16Updated this week
- Generating Image Specific Text☆29Aug 14, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆126Feb 21, 2023Updated 3 years ago
- ☆92Apr 15, 2022Updated 4 years ago
- ☆10Aug 26, 2022Updated 4 years ago
- Dual-Branch Network for Portrait Image Quality Assessment☆19Updated this week
- This repository contains the Adverbs in Recipes (AIR) dataset and the code published at the CVPR 23 paper: "Learning Action Changes by Me…☆13May 25, 2023Updated 3 years ago
- This is the pytorch implementation of the UAI2023 paper "A Trajectory is Worth Three Sentences: Multimodal Transformer for Offline Reinf…☆11Oct 9, 2023Updated 2 years ago
- ICLR 2023 DeCap: Decoding CLIP Latents for Zero-shot Captioning☆144Mar 16, 2023Updated 3 years ago