mesnico/TERAN

Readme badge preview -

If you own this repo, copy the snippet below and add it to your README.md

[![RelatedRepos](https://img.shields.io/badge/related-repos-yellow)](https://relatedrepos.com/gh/mesnico/TERAN)

mesnico / TERAN

Code and Resources for the Transformer Encoder Reasoning and Alignment Network (TERAN), accepted for publication in ACM Transactions on Multimedia Computing, Communications, and Applications (TOMM)

☆74

Alternatives and similar repositories for TERAN

Users that are interested in TERAN are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.

Sorting:

mesnico / TERN
View on GitHub
Code and Resources for the Transformer Encoder Reasoning Network (TERN) - https://arxiv.org/abs/2004.09144
☆58Dec 6, 2023Updated 2 years ago
yiling2018 / saem
View on GitHub
Learning Fragment Self-Attention Embeddings for Image-Text Matching, in ACM MM 2019
☆41Sep 24, 2019Updated 6 years ago
96-Zachary / vse_2ad
View on GitHub
☆15Apr 30, 2022Updated 4 years ago
mesnico / ALADIN
View on GitHub
Official implementation of the paper "ALADIN: Distilling Fine-grained Alignment Scores for Efficient Image-Text Matching and Retrieval"
☆28Dec 6, 2023Updated 2 years ago
sunnychencool / AOQ
View on GitHub
Adaptive Offline Quintuplet Loss for Image-Text Matching (AOQ)
☆34Jul 2, 2020Updated 6 years ago
Wordpress hosting with auto-scaling - Free Trial Offer • Ad
Fully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
woodfrog / vse_infty
View on GitHub
Code for "Learning the Best Pooling Strategy for Visual Semantic Embedding", CVPR 2021 (Oral)
☆165Aug 24, 2025Updated 11 months ago
iLearn-Lab / SIGIR21-DIME
View on GitHub
Dynamic Modality Interaction Modeling for Image-Text Retrieval. SIGIR'21
☆68Apr 5, 2026Updated 3 months ago
Paranioar / SGRAF
View on GitHub
[AAAI2021] The code of “Similarity Reasoning and Filtration for Image-Text Matching”
☆219Apr 11, 2024Updated 2 years ago
CrossmodalGroup / NAAF
View on GitHub
Implementation of our CVPR2022 paper, Negative-Aware Attention Framework for Image-Text Matching.
☆119Jun 19, 2023Updated 3 years ago
hardyqr / HAL
View on GitHub
[AAAI'20] Code release for "HAL: Improved Text-Image Matching by Mitigating Visual Semantic Hubs".
☆38Oct 4, 2023Updated 2 years ago
HuiChen24 / IMRAM
View on GitHub
code for our CVPR2020 paper "IMRAM: Iterative Matching with Recurrent Attention Memory for Cross-Modal Image-Text Retrieval"
☆95Mar 8, 2020Updated 6 years ago
fartashf / vsepp
View on GitHub
PyTorch Code for the paper "VSE++: Improving Visual-Semantic Embeddings with Hard Negatives"
☆523Dec 8, 2021Updated 4 years ago
BruceW91 / CVSE
View on GitHub
The official source code for the paper Consensus-Aware Visual-Semantic Embedding for Image-Text Matching (ECCV 2020)
☆168Feb 7, 2022Updated 4 years ago
kuanghuei / SCAN
View on GitHub
PyTorch source code for "Stacked Cross Attention for Image-Text Matching" (ECCV 2018)
☆579May 18, 2023Updated 3 years ago
AI Agents on DigitalOcean Gradient AI Platform • Ad
Build production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
LCFractal / TGDT
View on GitHub
Efficient Token-Guided Image-Text Retrieval with Consistent Multimodal Contrastive Training
☆30Jun 20, 2023Updated 3 years ago
fabiocarrara / features-adversarial-det
View on GitHub
Detect adversarial images from intermediate features in distance space
☆12Aug 22, 2018Updated 7 years ago
KunpengLi1994 / VSRN
View on GitHub
PyTorch code for ICCV'19 paper "Visual Semantic Reasoning for Image-Text Matching"
☆304Jan 14, 2020Updated 6 years ago
cwj1412 / MSCOCO-Flikcr30K_FG
View on GitHub
Benchmark data for "Rethinking Benchmarks for Cross-modal Image-text Retrieval" (SIGIR 2023)
☆28Apr 24, 2023Updated 3 years ago
mesnico / learning-relationship-aware-visual-features
View on GitHub
Relational Content-Based Image Retrieval (R-CBIR) - Retrieving images with given relationships among objects
☆17Oct 12, 2021Updated 4 years ago
IRVLUTD / HoloLens2ResearchTools
View on GitHub
The research tools developed for HoloLens2
☆10Dec 2, 2022Updated 3 years ago
Wangt-CN / MTFN-RR-PyTorch-Code
View on GitHub
The offical code for paper "Matching Images and Text with Multi-modal Tensor Fusion and Re-ranking", ACM Multimedia 2019 Oral
☆67Sep 28, 2019Updated 6 years ago
PKU-ICST-MIPL / MKVSE-TOMM2023
View on GitHub
☆28May 16, 2023Updated 3 years ago
CrossmodalGroup / GSMN
View on GitHub
Implementation of our CVPR2020 paper, Graph Structured Network for Image-Text Matching
☆170Oct 12, 2020Updated 5 years ago
Managed hosting for WordPress and PHP on Cloudways • Ad
Managed hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
Paranioar / Awesome_Matching_Pretraining_Transfering
View on GitHub
The Paper List of Large Multi-Modality Model (Perception, Generation, Unification), Parameter-Efficient Finetuning, Vision-Language Pretr…
☆446Sep 25, 2025Updated 10 months ago
cyh-sj / CGMN
View on GitHub
The code of the paper "Cross-Modal Graph Matching Network for Image-Text Retrieval" in ACM Transactions on Multimedia Computing, Communic…
☆45Jun 5, 2023Updated 3 years ago
YingZhangDUT / Cross-Modal-Projection-Learning
View on GitHub
TensorFlow Implementation of Deep Cross-Modal Projection Learning
☆95Nov 7, 2019Updated 6 years ago
biomedia-mira / cxr-foundation-bias
View on GitHub
Official repository for 'Risk of Bias in Chest Radiography Deep Learning Foundation Models'
☆12Sep 27, 2023Updated 2 years ago
jwehrmann / retrieval.pytorch
View on GitHub
Adaptive Cross-Modal Embeddings for Image-Sentence Alignment
☆36Oct 3, 2023Updated 2 years ago
HaoYang0123 / Position-Focused-Attention-Network
View on GitHub
Position Focused Attention Network for Image-Text Matching
☆69Aug 20, 2019Updated 6 years ago
LgQu / CAMERA
View on GitHub
Context-Aware Multi-View Summarization Network for Image-Text Matching. ACM MM'20
☆29May 26, 2022Updated 4 years ago
yalesong / pvse
View on GitHub
Polysemous Visual-Semantic Embedding for Cross-Modal Retrieval (CVPR 2019)
☆135Mar 15, 2024Updated 2 years ago
ZhangXu0963 / VSL
View on GitHub
The code of "Image-text Retrieval via Preserving Main Semantic of Vision" in ICME 2023.
☆15Dec 25, 2023Updated 2 years ago
GPUs on demand by Runpod - Special Offer Available • Ad
Run AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
kdwonn / DivE
View on GitHub
Repository of "Improving Cross-Modal Retrieval With Set of Diverse Embeddings" (CVPR'23, Highlight)
☆41Nov 15, 2023Updated 2 years ago
CrossmodalGroup / ESL
View on GitHub
☆12May 3, 2024Updated 2 years ago
hardyqr / Visual-Semantic-Embeddings-an-incomplete-list
View on GitHub
A paper list of visual semantic embeddings and text-image retrieval.
☆41Dec 4, 2020Updated 5 years ago
cdluminate / ladderloss
View on GitHub
Ladder Loss for Coherent Visual-Semantic Embedding, AAAI, 2020
☆13Aug 14, 2021Updated 4 years ago
DaniloSorano / PassNet
View on GitHub
A Computer Vision Approach for Pass Detection on Soccer Broadcast Video
☆22May 8, 2023Updated 3 years ago
peteanderson80 / bottom-up-attention
View on GitHub
Bottom-up attention model for image captioning and VQA, based on Faster R-CNN and Visual Genome
☆1,470Feb 3, 2023Updated 3 years ago
google-research-datasets / Crisscrossed-Captions
View on GitHub
Extended Intramodal and Intermodal Semantic Similarity Judgments for MS-COCO
☆54Sep 3, 2020Updated 5 years ago