Segment Anything with Deictic Prompting
☆27May 13, 2025Updated last year
Alternatives and similar repositories for deictic-segment-anything
Users that are interested in deictic-segment-anything are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official Repo for DAC-RL: Training LLMs for Divide-and-Conquer Reasoning Elevates Test-Time Scalability☆20Updated this week
- ☆10Apr 7, 2025Updated last year
- Code Release for ECCV 2024, "PCF-Lift: Panoptic Lifting by Probabilistic Contrastive Fusion"☆21Mar 23, 2025Updated last year
- Official PyTorch implementation of RACRO (https://www.arxiv.org/abs/2506.04559)☆19Jul 1, 2025Updated last year
- This is a repository contains the implementation of our NeurIPS'24 paper "Temporal Sentence Grounding with Relevance Feedback in Videos"☆13Aug 22, 2025Updated last year
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- [ECCV 2024] The first zero-shot setting for spatio-temporal video grounding.☆11Jul 16, 2024Updated 2 years ago
- [ICLR2023] Video Scene Graph Generation from Single-Frame Weak Supervision☆12Sep 17, 2023Updated 3 years ago
- Implementation of ''VPUFormer: Visual Prompt Unified Transformer for Interactive Image Segmentation''☆15Sep 16, 2025Updated last year
- [CVPR 2024] LoSh: Long-Short Text Joint Prediction Network for Referring Video Object Segmentation☆13Jun 17, 2024Updated 2 years ago
- [ACM MM-2024] RefMask3D: Language-Guided Transformer for 3D Referring Segmentation☆65Jul 29, 2024Updated 2 years ago
- View planning with multi-turn VLM agents: ViewSuite 6-DoF benchmark on real ScanNet scenes + iterative RL-SFT training☆25Sep 6, 2026Updated 2 weeks ago
- ☆34Feb 29, 2024Updated 2 years ago
- [ICCV 2023 Workshop] The Official Implementation of The First Prize Solution for RVOS Competition☆14Jan 1, 2024Updated 2 years ago
- Repository for "Rescan: Inductive Instance Segmentation for Indoor RGBD Scans" (ICCV 2019)☆17Mar 12, 2020Updated 6 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ALTo: Adaptive-Length Tokenizer for Autoregressive Mask Generation☆29May 27, 2025Updated last year
- Sa2VA-i is an improved version of the popular Sa2VA model☆17Nov 25, 2025Updated 9 months ago
- [ECCV 2026] Real-Time Interactive Multi-Target Video Segmentation☆68Jul 10, 2026Updated 2 months ago
- [NeurIPS 2024] A Unified Framework for 3D Scene Understanding☆180Jul 7, 2025Updated last year
- Open-Vocabulary SAM3D: Understand Any 3D Scene☆44Jun 9, 2025Updated last year
- The offical implement of ImbSAM (Imbalanced-SAM)☆28Mar 4, 2024Updated 2 years ago
- Strefer: Empowering Video LLMs with Space-Time Referring and Reasoning via Synthetic Instruction Data☆19Jun 2, 2026Updated 3 months ago
- Referring Image Segmentation Benchmarking with Segment Anything Model (SAM)☆39Apr 7, 2023Updated 3 years ago
- Online video temporal grounding☆16Oct 20, 2025Updated 11 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ECCV 2024 Oral] ActionVOS: Actions as Prompts for Video Object Segmentation☆32Dec 4, 2024Updated last year
- ☆13Oct 30, 2023Updated 2 years ago
- [ICCV 2025] GroundingSuite: Measuring Complex Multi-Granular Pixel Grounding☆77Jun 26, 2025Updated last year
- [ECCV'24] OpenIns3D: Snap and Lookup for 3D Open-vocabulary Instance Segmentation☆210Oct 19, 2024Updated last year
- The official repo for "Stepping Stones: A Progressive Training Strategy for Audio-Visual Semantic Segmentation", ECCV 2024☆18Oct 11, 2024Updated last year
- [NeurIPS2023] Code release for "Hierarchical Open-vocabulary Universal Image Segmentation"☆293Jun 19, 2025Updated last year
- [ACMMM 2025] Officially implement of the paper "Seg-Wild: Interactive Segmentation based on 3D Gaussian Splatting for Unconstrained Image…☆22Jul 29, 2025Updated last year
- Video-R2: Reinforcing Consistent and Grounded Reasoning in Multimodal Language Models☆19Sep 5, 2026Updated 2 weeks ago
- This is an official pytorch implementation of 'Group-wise Inhibition based Feature Regularization for Robust Classification' (ICCV 2021 a…☆10Dec 10, 2022Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICCV 2023] OnlineRefer: A Simple Online Baseline for Referring Video Object Segmentation☆58Oct 7, 2023Updated 2 years ago
- ☆148Jan 11, 2024Updated 2 years ago
- ICCV23 "Householder Projector for Unsupervised Latent Semantics Discovery"☆16Sep 7, 2026Updated 2 weeks ago
- Kaggle TGS Salt Identification Challenge (top 1%)☆15Nov 28, 2018Updated 7 years ago
- This is the official implementation of RGNet: A Unified Retrieval and Grounding Network for Long Videos☆20Mar 3, 2025Updated last year
- A community-driven AI automation framework that builds upon the incredible work of the open source community. Our goal is to combine lang…☆14Mar 23, 2025Updated last year
- SpaceVLLM: Endowing Multimodal Large Language Model with Spatio-Temporal Video Grounding Capability☆17May 8, 2025Updated last year