Cross-Modal Self-Attention Network for Referring Image Segmentation cvpr19
☆57Sep 11, 2019Updated 7 years ago
Alternatives and similar repositories for CMSA-Net
Users that are interested in CMSA-Net are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for Linguistic Structure Guided Context Modeling for Referring Image Segmentation, ECCV2020.☆16Oct 2, 2020Updated 5 years ago
- Referring Expression Object Segmentation with Caption-Aware Consistency, BMVC 2019☆31Apr 21, 2021Updated 5 years ago
- ☆47Oct 3, 2023Updated 2 years ago
- ☆38Jul 23, 2017Updated 9 years ago
- Dynamic Multimodal Instance Segmentation Guided by Natural Language Queries, ECCV 2018☆76Sep 21, 2021Updated 5 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ICME'22] Visual Grounding with Transformers☆28May 27, 2022Updated 4 years ago
- Referring Expression Parser☆27Feb 10, 2018Updated 8 years ago
- Inferring and Executing Programs for Visual Reasoning☆21Jan 4, 2019Updated 7 years ago
- SalNet on Keras: A deep convolutional network for saliency prediction☆11Jun 23, 2017Updated 9 years ago
- ☆14Jul 13, 2021Updated 5 years ago
- RefVOS☆28Feb 3, 2021Updated 5 years ago
- MAttNet: Modular Attention Network for Referring Expression Comprehension☆299Nov 29, 2022Updated 3 years ago
- A Diagnostic Dataset for Compositional Language and Elementary Visual Reasoning☆27Jan 20, 2022Updated 4 years ago
- [CVPR2020] Multi-task Collaborative Network for Joint Referring Expression Comprehension and Segmentation, CVPR2020 (oral)☆139Aug 4, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- An unofficial pytorch implementation of "TransVG: End-to-End Visual Grounding with Transformers".☆50Jun 7, 2021Updated 5 years ago
- The toolbox for the Google Refexp dataset proposed in this paper: http://arxiv.org/abs/1511.02283☆166Mar 1, 2017Updated 9 years ago
- Dataset API for "PhraseCut: Language-based Image Segmentation in the Wild"☆115Mar 28, 2026Updated 5 months ago
- Evaluation Framework for DAVIS 2017 Semi-supervised and Unsupervised used in the DAVIS Challenges☆201Feb 26, 2023Updated 3 years ago
- ☆235Apr 13, 2023Updated 3 years ago
- [ICCV2021 & TPAMI2023] Vision-Language Transformer and Query Generation for Referring Segmentation☆367Jan 7, 2022Updated 4 years ago
- Refer-Youtube-VOS dataset☆27Mar 10, 2026Updated 6 months ago
- (TIP 2024) Towards Robust Referring Image Segmentation☆41Mar 2, 2024Updated 2 years ago
- End-to-end Multi-modal Video Temporal Grounding, NeurIPS 2021☆18Oct 24, 2021Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Deep Active Contour Network for Medical Image Segmentation☆20Nov 16, 2020Updated 5 years ago
- Simple vs complex temporal recurrences for video saliency prediction (BMVC 2019)☆27Nov 22, 2022Updated 3 years ago
- A collection of papers about Referring Image Segmentation.☆830Jan 28, 2026Updated 7 months ago
- ☆12Oct 21, 2019Updated 6 years ago
- [CVPR2021] Look before you leap: learning landmark features for one-stage visual grounding.☆51Aug 31, 2021Updated 5 years ago
- Phrase Localization Evaluation Toolkit☆20Aug 16, 2019Updated 7 years ago
- ☆13Jan 19, 2024Updated 2 years ago
- S3 automatic lossless image compression☆10Aug 28, 2015Updated 11 years ago
- Training code for "SSTVOS: Sparse Spatiotemporal Transformers for Video Object Segmentation"☆88Nov 21, 2021Updated 4 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- This is the repo for Multi-level textual grounding☆34Jul 21, 2020Updated 6 years ago
- Code for CVPR 2021 paper. "Calibrated RGB-D Salient Object detection".☆45Oct 20, 2021Updated 4 years ago
- Gender/Age attribute grounding using weak supervised manner.☆12Jun 23, 2019Updated 7 years ago
- ☆12Jul 18, 2018Updated 8 years ago
- Code for the AAAI 2021 paper "Attributes-Guided and Pure-Visual Attention Alignment for Few-Shot Recognition".☆10Nov 21, 2022Updated 3 years ago
- The speaker-labeled information of LRW dataset, which is the outcome of the paper "Speaker-adaptive Lip Reading with User-dependent Paddi…☆10Oct 12, 2023Updated 2 years ago
- This repository is an official PyTorch implementation of our paper "Feature Distillation Interaction Weighting Network for Lightweight Im…☆21May 10, 2023Updated 3 years ago