Official Implementation for paper "Referring Transformer: A One-step Approach to Multi-task Visual Grounding" Neurips 2021
☆67May 26, 2022Updated 4 years ago
Alternatives and similar repositories for RefTR
Users that are interested in RefTR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Improving Visual Grounding with Visual-Linguistic Verification and Iterative Reasoning, CVPR 2022☆98Dec 2, 2022Updated 3 years ago
- Encoder Fusion Network with Co-Attention Embedding for Referring Image Segmentation, CVPR2021☆21Aug 17, 2021Updated 5 years ago
- ☆234Apr 13, 2023Updated 3 years ago
- ☆199Feb 27, 2024Updated 2 years ago
- A lightweight codebase for referring expression comprehension and segmentation☆57May 21, 2022Updated 4 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆92Apr 15, 2022Updated 4 years ago
- ☆42Jun 3, 2022Updated 4 years ago
- IROS 2023 "VL-Grasp: a 6-Dof Interactive Grasp Policy for Language-Oriented Objects in Cluttered Indoor Scenes"☆63Apr 22, 2024Updated 2 years ago
- MAttNet: Modular Attention Network for Referring Expression Comprehension☆299Nov 29, 2022Updated 3 years ago
- SeqTR: A Simple yet Universal Network for Visual Grounding☆144Oct 30, 2024Updated last year
- Improving One-stage Visual Grounding by Recursive Sub-query Construction, ECCV 2020☆90Sep 30, 2021Updated 4 years ago
- SOIT: Segmenting Objects with Instance-Aware Transformers☆14Jun 6, 2022Updated 4 years ago
- ☆10Jan 9, 2025Updated last year
- [CVPR2020] Multi-task Collaborative Network for Joint Referring Expression Comprehension and Segmentation, CVPR2020 (oral)☆139Aug 4, 2022Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- awesome grounding: A curated list of research papers in visual grounding☆1,127Sep 21, 2025Updated 11 months ago
- [ICCV2021 & TPAMI2023] Vision-Language Transformer and Query Generation for Referring Segmentation☆367Jan 7, 2022Updated 4 years ago
- A curated list of research papers in Referring Expression Comprehension (REC)☆46May 13, 2021Updated 5 years ago
- A Fast and Accurate One-Stage Approach to Visual Grounding, ICCV 2019 (Oral)☆149Nov 18, 2020Updated 5 years ago
- A benchmark dataset for GREx: GRES, GREC, and GREG [CVPR 2023 & IJCV 2026]☆241Nov 14, 2025Updated 9 months ago
- Lightweight Transformer for Multi-modal Tasks☆16Dec 9, 2022Updated 3 years ago
- A collection of papers about Referring Image Segmentation.☆828Jan 28, 2026Updated 7 months ago
- Learning phrase grounding from captioned images through InfoNCE bound on mutual information☆73Aug 22, 2020Updated 6 years ago
- Referring Expression Parser☆27Feb 10, 2018Updated 8 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [CVPR2022] Official Implementation of ReferFormer☆356Feb 15, 2025Updated last year
- [CVPR-2023] The official dataset of Advancing Visual Grounding with Scene Knowledge: Benchmark and Method.☆34Jul 12, 2023Updated 3 years ago
- [ACM MM 22] Correspondence Matters for Video Referring Expression Comprehension☆15Sep 4, 2022Updated 3 years ago
- [CVPR2021] Look before you leap: learning landmark features for one-stage visual grounding.☆51Aug 31, 2021Updated 5 years ago
- ☆13Oct 30, 2023Updated 2 years ago
- ☆27Oct 7, 2021Updated 4 years ago
- ☆14Nov 4, 2022Updated 3 years ago
- [ICRA 2025] A Parameter-Efficient Tuning Framework for Language-guided Object Grounding and Robot Grasping☆13Feb 7, 2025Updated last year
- Referring Expression Datasets API☆576Aug 27, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [AAAI 2023] DQ-DETR: Dual Query Detection Transformer for Phrase Extraction and Grounding☆58Nov 28, 2022Updated 3 years ago
- Official code for the paper, "TaCA: Upgrading Your Visual Foundation Model with Task-agnostic Compatible Adapter".☆16Jun 20, 2023Updated 3 years ago
- ☆17Nov 14, 2022Updated 3 years ago
- ☆1,052Oct 3, 2022Updated 3 years ago
- An official PyTorch implementation of the CRIS paper☆282Jun 9, 2024Updated 2 years ago
- ☆47Oct 3, 2023Updated 2 years ago
- Source code for EMNLP 2022 paper “PEVL: Position-enhanced Pre-training and Prompt Tuning for Vision-language Models”☆49Nov 10, 2022Updated 3 years ago