hardyqr/Visual-Semantic-Embeddings-an-incomplete-list

Readme badge preview -

If you own this repo, copy the snippet below and add it to your README.md

[![RelatedRepos](https://img.shields.io/badge/related-repos-yellow)](https://relatedrepos.com/gh/hardyqr/Visual-Semantic-Embeddings-an-incomplete-list)

hardyqr / Visual-Semantic-Embeddings-an-incomplete-list

A paper list of visual semantic embeddings and text-image retrieval.

☆41

Alternatives and similar repositories for Visual-Semantic-Embeddings-an-incomplete-list

Users that are interested in Visual-Semantic-Embeddings-an-incomplete-list are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.

Sorting:

sunnychencool / AOQ
View on GitHub
Adaptive Offline Quintuplet Loss for Image-Text Matching (AOQ)
☆34Jul 2, 2020Updated 6 years ago
yiling2018 / saem
View on GitHub
Learning Fragment Self-Attention Embeddings for Image-Text Matching, in ACM MM 2019
☆41Sep 24, 2019Updated 6 years ago
hardyqr / HAL
View on GitHub
[AAAI'20] Code release for "HAL: Improved Text-Image Matching by Mitigating Visual Semantic Hubs".
☆38Oct 4, 2023Updated 2 years ago
HuiChen24 / IMRAM
View on GitHub
code for our CVPR2020 paper "IMRAM: Iterative Matching with Recurrent Attention Memory for Cross-Modal Image-Text Retrieval"
☆95Mar 8, 2020Updated 6 years ago
ZihaoWang-CV / CAMP_iccv19
View on GitHub
CAMP: Cross-Modal Adaptive Message Passing for Text-Image Retrieval
☆126Feb 26, 2020Updated 6 years ago
Managed hosting for WordPress and PHP on Cloudways • Ad
Managed hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
Shiyang-Yan / Discrete-continous-PG-for-Retrieval
View on GitHub
☆13Feb 1, 2022Updated 4 years ago
BruceW91 / CVSE
View on GitHub
The official source code for the paper Consensus-Aware Visual-Semantic Embedding for Image-Text Matching (ECCV 2020)
☆168Feb 7, 2022Updated 4 years ago
kywen1119 / DSRAN
View on GitHub
Code for journal paper "Learning Dual Semantic Relations with Graph Attention for Image-Text Matching", TCSVT, 2020.
☆74Oct 25, 2022Updated 3 years ago
jwehrmann / retrieval.pytorch
View on GitHub
Adaptive Cross-Modal Embeddings for Image-Sentence Alignment
☆36Oct 3, 2023Updated 2 years ago
KunpengLi1994 / VSRN
View on GitHub
PyTorch code for ICCV'19 paper "Visual Semantic Reasoning for Image-Text Matching"
☆304Jan 14, 2020Updated 6 years ago
jwehrmann / lavse
View on GitHub
Language-Agnostic Visual-Semantic Embeddings (ICCV'19)
☆22Nov 11, 2019Updated 6 years ago
HaoYang0123 / Position-Focused-Attention-Network
View on GitHub
Position Focused Attention Network for Image-Text Matching
☆69Aug 20, 2019Updated 6 years ago
ItemZheng / KDDAug
View on GitHub
[ECCV2022] Rethinking Data Augmentation for Robust Visual Question Answering
☆13Nov 23, 2022Updated 3 years ago
yalesong / pvse
View on GitHub
Polysemous Visual-Semantic Embedding for Cross-Modal Retrieval (CVPR 2019)
☆135Mar 15, 2024Updated 2 years ago
Proton VPN Special Offer - Get 70% off • Ad
Special partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
princetonvisualai / SPICE-U
View on GitHub
☆11Sep 7, 2020Updated 5 years ago
CrossmodalGroup / GSMN
View on GitHub
Implementation of our CVPR2020 paper, Graph Structured Network for Image-Text Matching
☆170Oct 12, 2020Updated 5 years ago
m2man / LGSGM
View on GitHub
☆34Jun 14, 2022Updated 4 years ago
fortunechen / paper-reading_CrossModelGroup-USTC
View on GitHub
中科大跨模态智能组-每周论文分享
☆15Nov 20, 2022Updated 3 years ago
e-bug / volta
View on GitHub
[TACL 2021] Code and data for the framework in "Multimodal Pretraining Unmasked: A Meta-Analysis and a Unified Framework of Vision-and-La…
☆115Mar 24, 2022Updated 4 years ago
kuanghuei / SCAN
View on GitHub
PyTorch source code for "Stacked Cross Attention for Image-Text Matching" (ECCV 2018)
☆579May 18, 2023Updated 3 years ago
HuiChen24 / MM_SemanticConsistency
View on GitHub
code for our MM2019 paper “Cross-Modal Image-Text Retrieval with Semantic Consistency”
☆17Dec 7, 2019Updated 6 years ago
kkanshul / dock
View on GitHub
☆12Feb 14, 2019Updated 7 years ago
weijiawu / CisDQ
View on GitHub
☆13Nov 29, 2023Updated 2 years ago
GPUs on demand by Runpod - Special Offer Available • Ad
Run AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
fkxssaa / Deliberate-Attention-Networks-for-Image-Captioning
View on GitHub
Deliberate Attention Networks for Image Captioning (AAAI 2019)
☆11Sep 30, 2019Updated 6 years ago
cyh-sj / CGMN
View on GitHub
The code of the paper "Cross-Modal Graph Matching Network for Image-Text Retrieval" in ACM Transactions on Multimedia Computing, Communic…
☆45Jun 5, 2023Updated 3 years ago
mhelhoseiny / CIZSL
View on GitHub
Creativity Inspired Zero-Shot Learning
☆32Mar 8, 2021Updated 5 years ago
woodfrog / vse_infty
View on GitHub
Code for "Learning the Best Pooling Strategy for Visual Semantic Embedding", CVPR 2021 (Oral)
☆165Aug 24, 2025Updated 11 months ago
zjy526223908 / BTDA
View on GitHub
implement for paper Rethinking Domain Adaptation Blending target Domain Adaptation by Adversarial Meta Adaptation Network
☆38Mar 12, 2019Updated 7 years ago
mancinimassimiliano / adagraph
View on GitHub
PyTorch implementation of AdaGraph: Unifying Predictive and Continuous Domain Adaptation through Graphs
☆23Feb 6, 2023Updated 3 years ago
yue-zhongqi / tcm
View on GitHub
☆34Jul 28, 2021Updated 5 years ago
Wangt-CN / MTFN-RR-PyTorch-Code
View on GitHub
The offical code for paper "Matching Images and Text with Multi-modal Tensor Fusion and Re-ranking", ACM Multimedia 2019 Oral
☆67Sep 28, 2019Updated 6 years ago
UNITES-Lab / AgentSymbiotic
View on GitHub
☆14Mar 11, 2025Updated last year
Managed hosting for WordPress and PHP on Cloudways • Ad
Managed hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
niluthpol / multimodal_vtt
View on GitHub
Joint Embedding with Multimodal Cues for Cross-Modal Video-Text Retrieval
☆68Apr 10, 2020Updated 6 years ago
Dawn-LX / VidVRD-tracklets
View on GitHub
Video Visual Relation Detection (VidVRD) tracklets generation. also for ACM MM Visual Relation Understanding Grand Challenge
☆40Dec 5, 2022Updated 3 years ago
yangxuntu / SGAE
View on GitHub
☆218Feb 26, 2022Updated 4 years ago
iLearn-Lab / SIGIR21-DIME
View on GitHub
Dynamic Modality Interaction Modeling for Image-Text Retrieval. SIGIR'21
☆68Apr 5, 2026Updated 3 months ago
mesnico / TERAN
View on GitHub
Code and Resources for the Transformer Encoder Reasoning and Alignment Network (TERAN), accepted for publication in ACM Transactions on M…
☆74Dec 6, 2023Updated 2 years ago
ChopinSharp / ref-nms
View on GitHub
Official codebase for "Ref-NMS: Breaking Proposal Bottlenecks in Two-Stage Referring Expression Grounding"
☆22Dec 20, 2020Updated 5 years ago
christophschuhmann / 4MC-4M-Image-Text-Pairs-with-CLIP-embeddings
View on GitHub
I have created a dataset of Image-Text-Pairs by using the cosine similarity of the CLIP embeddings of the image & it's caption derrived f…
☆17Apr 22, 2021Updated 5 years ago