mesnico/TERN

Readme badge preview -

If you own this repo, copy the snippet below and add it to your README.md

[![RelatedRepos](https://img.shields.io/badge/related-repos-yellow)](https://relatedrepos.com/gh/mesnico/TERN)

mesnico / TERN

Code and Resources for the Transformer Encoder Reasoning Network (TERN) - https://arxiv.org/abs/2004.09144

☆58

Alternatives and similar repositories for TERN

Users that are interested in TERN are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.

Sorting:

mesnico / TERAN
View on GitHub
Code and Resources for the Transformer Encoder Reasoning and Alignment Network (TERAN), accepted for publication in ACM Transactions on M…
☆74Dec 6, 2023Updated 2 years ago
yiling2018 / saem
View on GitHub
Learning Fragment Self-Attention Embeddings for Image-Text Matching, in ACM MM 2019
☆41Sep 24, 2019Updated 6 years ago
sunnychencool / AOQ
View on GitHub
Adaptive Offline Quintuplet Loss for Image-Text Matching (AOQ)
☆34Jul 2, 2020Updated 6 years ago
kywen1119 / DSRAN
View on GitHub
Code for journal paper "Learning Dual Semantic Relations with Graph Attention for Image-Text Matching", TCSVT, 2020.
☆74Oct 25, 2022Updated 3 years ago
AndresPMD / semantic_adaptive_margin
View on GitHub
WACV 2022 Paper - Is An Image Worth Five Sentences? A New Look into Semantics for Image-Text Matching
☆16Dec 10, 2021Updated 4 years ago
Bare Metal GPUs on DigitalOcean Gradient AI • Ad
Purpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
Wangt-CN / MTFN-RR-PyTorch-Code
View on GitHub
The offical code for paper "Matching Images and Text with Multi-modal Tensor Fusion and Re-ranking", ACM Multimedia 2019 Oral
☆68Sep 28, 2019Updated 6 years ago
HuiChen24 / IMRAM
View on GitHub
code for our CVPR2020 paper "IMRAM: Iterative Matching with Recurrent Attention Memory for Cross-Modal Image-Text Retrieval"
☆95Mar 8, 2020Updated 6 years ago
ciampluca / Virtual-to-Real-Pedestrian-Detection
View on GitHub
Code and Resources for our paper "Virtual to Real adaptation of Pedestrian Detectors" - https://www.mdpi.com/1424-8220/20/18/5250
☆11Jul 25, 2024Updated last year
BruceW91 / CVSE
View on GitHub
The official source code for the paper Consensus-Aware Visual-Semantic Embedding for Image-Text Matching (ECCV 2020)
☆168Feb 7, 2022Updated 4 years ago
mesnico / learning-relationship-aware-visual-features
View on GitHub
Relational Content-Based Image Retrieval (R-CBIR) - Retrieving images with given relationships among objects
☆17Oct 12, 2021Updated 4 years ago
fawazsammani / show-edit-tell
View on GitHub
Show, Edit and Tell: A Framework for Editing Image Captions, CVPR 2020
☆82Jul 17, 2020Updated 6 years ago
mesnico / ALADIN
View on GitHub
Official implementation of the paper "ALADIN: Distilling Fine-grained Alignment Scores for Efficient Image-Text Matching and Retrieval"
☆28Dec 6, 2023Updated 2 years ago
ZihaoWang-CV / CAMP_iccv19
View on GitHub
CAMP: Cross-Modal Adaptive Message Passing for Text-Image Retrieval
☆127Feb 26, 2020Updated 6 years ago
KunpengLi1994 / VSRN
View on GitHub
PyTorch code for ICCV'19 paper "Visual Semantic Reasoning for Image-Text Matching"
☆304Jan 14, 2020Updated 6 years ago
Simple, predictable pricing with DigitalOcean hosting • Ad
Always know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
hardyqr / HAL
View on GitHub
[AAAI'20] Code release for "HAL: Improved Text-Image Matching by Mitigating Visual Semantic Hubs".
☆38Oct 4, 2023Updated 2 years ago
iLearn-Lab / SIGIR21-DIME
View on GitHub
Dynamic Modality Interaction Modeling for Image-Text Retrieval. SIGIR'21
☆70Apr 5, 2026Updated 3 months ago
HaoYang0123 / Position-Focused-Attention-Network
View on GitHub
Position Focused Attention Network for Image-Text Matching
☆69Aug 20, 2019Updated 6 years ago
ciampluca / PrACo
View on GitHub
☆16May 19, 2026Updated 2 months ago
CrossmodalGroup / GSMN
View on GitHub
Implementation of our CVPR2020 paper, Graph Structured Network for Image-Text Matching
☆170Oct 12, 2020Updated 5 years ago
CrossmodalGroup / BFAN
View on GitHub
Implementation of our ACMMM2019 paper, Focus Your Attention: A Bidirectional Focal Attention Network for Image-Text Matching
☆39Jun 19, 2023Updated 3 years ago
LivXue / GNN4CMR
View on GitHub
PyTorch implementation of the AAAI-21 paper "Dual Adversarial Label-aware Graph Neural Networks for Cross-modal Retrieval" and the TPAMI-…
☆42Nov 1, 2022Updated 3 years ago
fabiocarrara / meye
View on GitHub
A deep-learning-based web tool for translational and real-time pupillometry
☆54Jan 8, 2026Updated 6 months ago
woodfrog / vse_infty
View on GitHub
Code for "Learning the Best Pooling Strategy for Visual Semantic Embedding", CVPR 2021 (Oral)
☆165Aug 24, 2025Updated 10 months ago
Open source password manager - Proton Pass • Ad
Securely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
malihealikhani / Cross-modal_Coherence_Modeling
View on GitHub
Cross-modal Coherence Modeling for Caption Generation
☆11Jul 24, 2020Updated 5 years ago
ezeli / InSentiCap_model
View on GitHub
A pytorch implementation of our paper Image Captioning with Inherent Sentiment (ICME 2021 Oral).
☆11Jul 18, 2022Updated 4 years ago
XLearning-SCU / 2021-CVPR-MRL
View on GitHub
Learning Cross-modal Retrieval with Noisy Labels (CVPR 2021, PyTorch Code)
☆13Apr 7, 2021Updated 5 years ago
cyrilou242 / learning-lightnr
View on GitHub
Generate multiple choice fill-in-the-blank questions from any article.
☆13Dec 8, 2022Updated 3 years ago
fartashf / vsepp
View on GitHub
PyTorch Code for the paper "VSE++: Improving Visual-Semantic Embeddings with Hard Negatives"
☆523Dec 8, 2021Updated 4 years ago
niluthpol / multimodal_vtt
View on GitHub
Joint Embedding with Multimodal Cues for Cross-Modal Video-Text Retrieval
☆68Apr 10, 2020Updated 6 years ago
CuthbertCai / Ask-Confirm
View on GitHub
Ask&Confirm: Active Detail Enriching for Cross-Modal Retrieval with Partial Query (ICCV2021)
☆20Dec 4, 2021Updated 4 years ago
cyh-sj / CGMN
View on GitHub
The code of the paper "Cross-Modal Graph Matching Network for Image-Text Retrieval" in ACM Transactions on Multimedia Computing, Communic…
☆45Jun 5, 2023Updated 3 years ago
AnnikaLindh / Diverse_and_Specific_Image_Captioning
View on GitHub
Unsupervised specificity-guided optimization of Image Captioning models to encourage meaningful diversity in the generated captions. Code…
☆13May 25, 2025Updated last year
1-Click AI Models by DigitalOcean Gradient • Ad
Deploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
kuanghuei / SCAN
View on GitHub
PyTorch source code for "Stacked Cross Attention for Image-Text Matching" (ECCV 2018)
☆579May 18, 2023Updated 3 years ago
CrossmodalGroup / CMCAN
View on GitHub
Implementation of our AAAI2022 paper, Show Your Faith: Cross-Modal Confidence-Aware Network for Image-Text Matching.
☆36Jun 16, 2023Updated 3 years ago
RiTUAL-MBZUAI / multimodal_NER
View on GitHub
"Can images help recognize entities? A study of the role of images for Multimodal NER" (W-NUT at EMNLP 2021)
☆21Nov 14, 2021Updated 4 years ago
UKPLab / MMT-Retrieval
View on GitHub
☆131Dec 10, 2022Updated 3 years ago
LooperXX / ManagerTower
View on GitHub
Code for ACL 2023 Oral Paper: ManagerTower: Aggregating the Insights of Uni-Modal Experts for Vision-Language Representation Learning
☆12Aug 23, 2025Updated 10 months ago
LiJiaBei-7 / rivrl
View on GitHub
Source code of our TCSVT'22 paper Reading-strategy Inspired Visual Representation Learning for Text-to-Video Retrieval
☆19Feb 13, 2022Updated 4 years ago
tongshoujie / MATCH-TUNING
View on GitHub
MATCH-TUNING
☆15Aug 6, 2022Updated 3 years ago