A vision-language tracking paper list, articles related to visual language tracking have been documented.
☆46Dec 15, 2024Updated last year
Alternatives and similar repositories for Awesome-Vision-Language-Tracking
Users that are interested in Awesome-Vision-Language-Tracking are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A visual object tracking paper list, articles related to visual object tracking have been documented.☆58Nov 6, 2024Updated last year
- A series of improved methods are used for visual tracking☆10Nov 29, 2025Updated 8 months ago
- Vision-Language based Visual Object Tracking☆36Jul 4, 2026Updated last month
- [ICASSP'25] Enhancing Vision-Language Tracking by Effectively Converting Textual Cues into Visual Cues☆19Dec 31, 2024Updated last year
- [ICCV'23] CiteTracker: Correlating Image and Text for Visual Tracking☆49Jun 20, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Can we make visual tracking systems align more closely with human visual perception?☆44Jul 13, 2026Updated 3 weeks ago
- ☆48Updated this week
- Official implementation of "SPMTrack: Spatio-Temporal Parameter-Efficient Fine-Tuning with Mixture of Experts for Scalable Visual Trackin…☆59Oct 19, 2025Updated 9 months ago
- LoRAT_pytracking: reproduction of [ECCV2024] LoRAT☆47Dec 9, 2024Updated last year
- ☆24Dec 2, 2025Updated 8 months ago
- [AAAI25] Asymmetric Siamese Network for Efficient Visual Tracking☆51Mar 4, 2025Updated last year
- [NeurIPS2024] - SimVG: A Simple Framework for Visual Grounding with Decoupled Multi-modal Fusion☆103Oct 29, 2025Updated 9 months ago
- Run SOTA Vision-Language Model Florence-2 on your data!☆15Mar 27, 2025Updated last year
- [TCSVT2025] AVLTrack: Dynamic Sparse Learning for Aerial Vision-Language Tracking☆23Mar 10, 2026Updated 5 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Automatically update arXiv papers about SOT & VLT, Multi-modal Learning, LLM and Video Understanding using Github Actions.☆48Aug 3, 2026Updated last week
- A continuously updated project to track the latest progress in the field of multi-modal object tracking. This project focuses solely on s…☆1,074Updated this week
- [ICCV2025] PropVG: End-to-End Proposal-Driven Visual Grounding with Multi-Granularity Discrimination☆32Oct 13, 2025Updated 9 months ago
- TransMDOT☆22Jan 8, 2024Updated 2 years ago
- time series data, pre-training, 1d signal, fusion☆16Mar 6, 2026Updated 5 months ago
- [NeurIPS'24] GoMatching: A Simple Baseline for Video Text Spotting via Long and Short Term Matching☆34May 29, 2025Updated last year
- XiHeFusion: First Chat LLM for Nuclear Fusion☆16Feb 11, 2026Updated 5 months ago
- Official Repository for the ACM MM 2024 paper "Navigating Beyond Instructions: Vision-and-Language Navigation in Obstructed Environments"☆16May 16, 2025Updated last year
- The official implementation of the ECCV 2024 paper "Empowering Embodied Visual Tracking with Visual Foundation Models and Offline RL"☆23Oct 17, 2025Updated 9 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [CVPR 2025, IJCV 2026] "A Distractor-Aware Memory for Visual Object Tracking with SAM2", "Distractor-Aware Memory-Based Visual Object Tra…☆491Apr 7, 2026Updated 4 months ago
- DisTime: Distribution-based Time Representation for Video Large Language Models.☆21Jul 10, 2025Updated last year
- Ranking-Based Siamese Visual Tracking(CVPR2022)☆36Jul 13, 2023Updated 3 years ago
- [CVPR 2025] Feature4X: Bridging Any Monocular Video to 4D Agentic AI with Versatile Gaussian Feature Fields☆41Oct 18, 2025Updated 9 months ago
- ☆13Jun 17, 2026Updated last month
- [CVPR 2023] Referring Multi-Object Tracking☆160Jul 2, 2024Updated 2 years ago
- TrackGPT: Track What You Need in Videos via Text Prompts☆25May 16, 2023Updated 3 years ago
- Document Haystacks: Vision-Language Reasoning Over Piles of 1000+ Documents, CVPR 2025☆26Jan 25, 2025Updated last year
- The source code of "Material-Guided Multi-View Fusion Network for Hyperspectral Object Tracking".☆18Mar 29, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Official Code for "MITracker: Multi-View Integration for Visual Object Tracking"☆137Jun 18, 2025Updated last year
- ☆18Jan 6, 2026Updated 7 months ago
- ☆115Dec 17, 2024Updated last year
- 🚀 Reasoning-based Multi-Object Tracking☆26Apr 30, 2026Updated 3 months ago
- Some methods of image enhancement☆12Nov 1, 2022Updated 3 years ago
- watermark video delogo☆11Nov 27, 2020Updated 5 years ago
- Code for ORAR Agent for Vision and Language Navigation on Touchdown and map2seq☆20Nov 3, 2023Updated 2 years ago