TrackGPT: Track What You Need in Videos via Text Prompts
☆25May 16, 2023Updated 3 years ago
Alternatives and similar repositories for TrackGPT
Users that are interested in TrackGPT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆25Dec 23, 2024Updated last year
- [ICASSP'25] Enhancing Vision-Language Tracking by Effectively Converting Textual Cues into Visual Cues☆19Dec 31, 2024Updated last year
- ☆18Feb 8, 2026Updated 6 months ago
- [NeurIPS 2024] Repository for the paper "OVT-B: A New Large-Scale Benchmark for Open-Vocabulary Multi-Object Tracking".☆29Nov 9, 2024Updated last year
- Official Implementation of ECCV2024 paper: SLAck☆29Sep 18, 2024Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆15May 21, 2026Updated 2 months ago
- [PRCV-2024] State Space Model based Frame-Event Tracking☆53Dec 6, 2025Updated 8 months ago
- This repository contains the implementation of SAM3 trackers.☆37Jun 30, 2026Updated last month
- [ECCV 2024] Beyond MOT: Semantic Multi-Object Tracking☆55Nov 19, 2024Updated last year
- ☆13Jul 15, 2024Updated 2 years ago
- Can we make visual tracking systems align more closely with human visual perception?☆44Jul 13, 2026Updated last month
- [NeurIPS'24] MemVLT: Vision-Language Tracking with Adaptive Memory-based Prompts☆19Oct 7, 2024Updated last year
- ☆21Jul 25, 2024Updated 2 years ago
- The public reproducible analysis code used for the gaze project☆11May 16, 2026Updated 2 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆26Jul 22, 2026Updated 3 weeks ago
- ☆24Dec 20, 2024Updated last year
- Official Repo for MageBench: Bridging Large Multimodal Models to Agents☆21Jan 8, 2025Updated last year
- Multi-Granularity Language-Guided Multi-Object Tracking☆26Nov 3, 2025Updated 9 months ago
- Vision-Language based Visual Object Tracking☆36Jul 4, 2026Updated last month
- Segment Anything with Deictic Prompting☆27May 13, 2025Updated last year
- VPTracker: Global Vision-Language Tracking via Visual Prompt and MLLM☆16Mar 10, 2026Updated 5 months ago
- R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning.☆66May 14, 2025Updated last year
- (NeurIPS 2023) Open-set visual object query search & localization in long-form videos☆26Feb 1, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Using distilled CLIP model to deploy the android device☆20Feb 28, 2023Updated 3 years ago
- OVTrack: Open-Vocabulary Multiple Object Tracking [CVPR 2023]☆115Oct 14, 2024Updated last year
- The official implementation for the paper [Towards Unified Token Learning for Vision-Language Tracking].☆24Dec 13, 2023Updated 2 years ago
- [AAAI 2025] AL-Ref-SAM 2: Unleashing the Temporal-Spatial Reasoning Capacity of GPT for Training-Free Audio and Language Referenced Video…☆94Dec 23, 2024Updated last year
- [NeurIPS 2024] VastTrack: Vast Category Visual Object Tracking☆77Sep 30, 2025Updated 10 months ago
- CVPR24☆72Aug 4, 2024Updated 2 years ago
- Decoupled Memory Selection for Multi-target Video Segmentation of SAM3☆62Jan 16, 2026Updated 6 months ago
- AI-powered scientific figure generator using LLM analysis and nano banana for publication-quality visualizations☆19Feb 7, 2026Updated 6 months ago
- [ICCV 2023 Workshop] The Official Implementation of The First Prize Solution for RVOS Competition☆14Jan 1, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆54Feb 8, 2025Updated last year
- [ICCV2023] Isomer: Isomerous Transformer for Zero-Shot Video Object Segmentation☆30Nov 21, 2023Updated 2 years ago
- [CVPR2024] Towards Generalizable Multi-Object Tracking☆34May 3, 2024Updated 2 years ago
- ☆17Oct 4, 2024Updated last year
- 【NeurIPS 2024】The implementation of LIVE: Learnable In-Context Vector for Visual Question Answering https://arxiv.org/abs/2406.13185☆23May 31, 2025Updated last year
- ☆25Jan 29, 2026Updated 6 months ago
- PyTorch implementation of "Efficient Motion Prompt Learning for Robust Visual Tracking" (ICML2025)☆31Dec 17, 2025Updated 7 months ago