All in One: Exploring Unified Vision-Language Tracking with Multi-Modal Alignment
☆21Feb 11, 2025Updated last year
Alternatives and similar repositories for All-in-One
Users that are interested in All-in-One are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The official implementation for the CVPR 2023 paper Joint Visual Grounding and Tracking with Natural Language Specification.☆78Jun 3, 2023Updated 3 years ago
- Source code of the paper: Overlapped Trajectory-Enhanced Visual Tracking☆11Sep 3, 2024Updated 2 years ago
- [TCSVT2025] AVLTrack: Dynamic Sparse Learning for Aerial Vision-Language Tracking☆23Mar 10, 2026Updated 6 months ago
- The official implementation for the paper [Towards Unified Token Learning for Vision-Language Tracking].☆24Dec 13, 2023Updated 2 years ago
- The official pytorch implementation of our AAAI 2024 paper "Unifying Visual and Vision-Language Tracking via Contrastive Learning"☆52Nov 4, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- This repository contains the implementation of SAM3 trackers.☆37Jun 30, 2026Updated 2 months ago
- [NeurIPS 2024] VastTrack: Vast Category Visual Object Tracking☆77Sep 30, 2025Updated 11 months ago
- Combining OSTrack and Segment Anything for VOT and VOS☆14Apr 10, 2023Updated 3 years ago
- Bi-directional Adapter for Multi-modal Tracking☆104Mar 19, 2024Updated 2 years ago
- ☆53Dec 23, 2022Updated 3 years ago
- ☆15Dec 3, 2021Updated 4 years ago
- [AAAI2025] SUTrack: Towards Simple and Unified Single Object Tracking☆169Jun 16, 2025Updated last year
- Local self-attention in Transformer for visual question answering☆13Mar 17, 2024Updated 2 years ago
- Vision-Language based Visual Object Tracking☆36Jul 4, 2026Updated 2 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆13Jul 15, 2024Updated 2 years ago
- PyTorch implementation of "Leveraging the Power of Data Augmentation for Transformer-based Tracking" (WACV2024)☆14Nov 14, 2023Updated 2 years ago
- Towards More Flexible and Accurate Object Tracking with Natural Language: Algorithms and Benchmark (CVPR 2021)☆56Mar 2, 2026Updated 6 months ago
- Paper list for vision-language tracking☆75Jul 1, 2026Updated 2 months ago
- TransMDOT☆22Jan 8, 2024Updated 2 years ago
- CVPR24☆73Aug 4, 2024Updated 2 years ago
- WebUAV-3M: A million-scale multi-modal UAV tracking benchmark☆75Apr 3, 2026Updated 5 months ago
- Code and models for RELO☆26May 16, 2026Updated 4 months ago
- Official implementation of "SPMTrack: Spatio-Temporal Parameter-Efficient Fine-Tuning with Mixture of Experts for Scalable Visual Trackin…☆62Oct 19, 2025Updated 11 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Multi-Granularity Language-Guided Multi-Object Tracking☆26Nov 3, 2025Updated 10 months ago
- The implement of "Learning Spatial-Frequency Transformer for Visual Object Tracking"☆20Jun 29, 2023Updated 3 years ago
- Official Implementation of Video-MA2MBA☆12Dec 3, 2024Updated last year
- ☆25Dec 23, 2024Updated last year
- Official implementation of the TransT-M (the winner of VOT-RT 2021) , including code and models.☆28Mar 28, 2023Updated 3 years ago
- Code for our paper "FocusTrack: A Self-Adaptive Local Sampling Algorithm for Efficient Anti-UAV Tracking"☆41Apr 21, 2025Updated last year
- Robust Tracking via Mamba-based Context-aware Token Learning (AAAI 2025)☆16Nov 6, 2025Updated 10 months ago
- M-SpecGene: Generalized Foundation Model for RGBT Multispectral Vision (ICCV 2025)☆41Nov 19, 2025Updated 10 months ago
- #ICCV 2025, # FlexTrack,# Mixture of Experts☆22Sep 5, 2026Updated 2 weeks ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [IEEE TMM 2025] CRSOT: Cross-Resolution Object Tracking using Unaligned Frame and Event Cameras☆22Jan 18, 2025Updated last year
- The official implementation for the paper [ODTrack: Online Dense Temporal Token Learning for Visual Tracking].☆192Oct 7, 2024Updated last year
- VPTracker: Global Vision-Language Tracking via Visual Prompt and MLLM☆17Sep 14, 2026Updated last week
- LoRAT_pytracking: reproduction of [ECCV2024] LoRAT☆46Dec 9, 2024Updated last year
- Modality-missing RGBT Tracking: Invertible Prompt Learning and High-quality Benchmarks (IJCV2024))☆27Mar 13, 2026Updated 6 months ago
- The official implementation for the CVPR'2025 paper Dynamic Updates for Language Adaptation in Visual-Language Tracking☆45Mar 27, 2025Updated last year
- ☆63Sep 25, 2025Updated 11 months ago