[TPAMI2025] Improving Generalized Visual Grounding with Instance-aware Joint Learning
☆33Apr 28, 2026Updated 4 months ago
Alternatives and similar repositories for InstanceVG
Users that are interested in InstanceVG are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV2025] PropVG: End-to-End Proposal-Driven Visual Grounding with Multi-Granularity Discrimination☆32Oct 13, 2025Updated 10 months ago
- [ECCV2026] MomentSeg: Moment-Centric Sampling for Enhanced Video Pixel Understanding☆25Jun 19, 2026Updated 2 months ago
- [AAAI2025 selected as oral] - Multi-task Visual Grounding with Coarse-to-Fine Consistency Constraints☆45Jul 2, 2025Updated last year
- [ICCV2025] DeRIS: Decoupling Perception and Cognition for Enhanced Referring Image Segmentation through Loopback Synergy☆48Nov 21, 2025Updated 9 months ago
- [PR2026] Drone Referring Localization: An Efficient Heterogeneous Spatial Feature Interaction Method For UAV Self-Localization☆96Feb 19, 2026Updated 6 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [NeurIPS2024] - SimVG: A Simple Framework for Visual Grounding with Decoupled Multi-modal Fusion☆103Oct 29, 2025Updated 10 months ago
- 「TCSVT2021」A Transformer-Based Feature Segmentation and Region Alignment Method For UAV-View Geo-Localization☆122Mar 7, 2024Updated 2 years ago
- ☆12Oct 24, 2023Updated 2 years ago
- This is the implementation of the paper "Test-Time Adaptive Object Detection with Foundation Model" (Neurips 2025)☆23Jan 30, 2026Updated 7 months ago
- Code release for "Strike a Balance in Continual Panoptic Segmentation" (ECCV 2024)☆14Mar 14, 2025Updated last year
- ☆31Sep 22, 2025Updated 11 months ago
- Official Implementation of "OVS Meets Continual Learning: Towards Sustainable Open-Vocabulary Segmentation" (NeurIPS 2025).☆16Feb 27, 2026Updated 6 months ago
- Target-Grounded Graph-Aware Transformer for Aerial Vision-and-Dialog Navigation, AVDN Challenge, ICCV CLVL 2023.☆21Jan 2, 2024Updated 2 years ago
- paper list on Video Moment Retrieval (VMR), or Temporal Video Grounding (TVG), Video Grounding (VG), or Temporal Sentence Grounding in Vi…☆43Jul 30, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- EagleVision: Object-level Attribute Multimodal LLM for Remote Sensing☆26May 29, 2025Updated last year
- [WACV 2026] Official implementation of the paper: “CountingDINO: A Training-free Pipeline for Exemplar-based Class-Agnostic Counting”☆64Jun 22, 2026Updated 2 months ago
- [ICCV 2025] MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation☆23Sep 5, 2025Updated 11 months ago
- LLaVA-Next for STVG☆21Dec 5, 2025Updated 8 months ago
- This is the open-sourced link of the TPAMI 2026 paper "SkyFind: A Large-Scale Benchmark Unveiling Referring Expression Comprehension for …☆33Aug 12, 2026Updated 2 weeks ago
- RefDrone: A Challenging Benchmark for Drone Scene Referring Expression Comprehension☆47Jul 8, 2026Updated last month
- [CVPR2024] Mask Grounding for Referring Image Segmentation☆29Jul 22, 2024Updated 2 years ago
- ☆28Feb 21, 2025Updated last year
- [CVPR2024] GSVA: Generalized Segmentation via Multimodal Large Language Models☆167Sep 12, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Emergent Visual Grounding in Large Multimodal Models Without Grounding Supervision☆47Oct 19, 2025Updated 10 months ago
- [ICCV 2023] ADNet: Lane Shape Prediction via Anchor Decomposition☆37Oct 11, 2023Updated 2 years ago
- When Pixel Difference Patterns Meet ViT: PiDiViT for Few-Shot Object Detection☆19Nov 3, 2025Updated 9 months ago
- An implementation of improved incremental Singular Value Decomposition(iSVD) algorithm☆17Jun 3, 2026Updated 2 months ago
- ☆29Apr 8, 2025Updated last year
- Improving One-stage Visual Grounding by Recursive Sub-query Construction, ECCV 2020☆90Sep 30, 2021Updated 4 years ago
- ViGiL3D: A Linguistically Diverse Dataset for 3D Visual Grounding☆20Aug 8, 2025Updated last year
- Incrementally compute an approximate truncated singular value decomposition☆20Jun 29, 2026Updated 2 months ago
- ☆29Apr 2, 2026Updated 4 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆21Sep 16, 2025Updated 11 months ago
- ☆10Jan 6, 2025Updated last year
- 😎 up-to-date & curated list of awesome 3D Visual Grounding papers, methods & resources.☆282Jan 14, 2026Updated 7 months ago
- Pytorch implementation of Each Part Matters: Local Patterns Facilitate Cross-view Geo-localization https://arxiv.org/abs/2008.11646☆99Jul 6, 2026Updated last month
- Proposed fuzzy reward model with GRPO to improve VLM's abilities in crowd counting task.☆21Apr 11, 2025Updated last year
- [AAAI 2026] Relation-R1: Progressively Cognitive Chain-of-Thought Guided Reinforcement Learning for Unified Relation Comprehension☆20Mar 6, 2026Updated 5 months ago
- (CVPR25) Exploring Contextual Attribute Density in Referring Expression Counting☆20Dec 3, 2025Updated 8 months ago