🚁 Can Vision-Language Models Think from the Sky? UAVReason for Aerial Reasoning and Generation
☆24Jul 11, 2026Updated 2 months ago
Alternatives and similar repositories for UAVReason
Users that are interested in UAVReason are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 📌 Official baseline implementation of 'Last-Meter Precision Navigation for UAVs: A Diffusion-Refined Aerial Visual Servoing Approach‘, s…☆54Jul 7, 2026Updated 2 months ago
- SEED Dataset☆29Jun 3, 2025Updated last year
- Official Repo for Look, Compare and Draw: Differential Query Transformer for Automatic Oil Painting☆17Mar 31, 2026Updated 5 months ago
- The official repo of "Towards Scalable Video Anomaly Retrieval: A Synthetic Video-Text Benchmark"☆21Jun 5, 2025Updated last year
- [CVPR 2026] Code for "The Coherence Trap: When MLLM-Crafted Narratives Exploit Manipulated Visual Contexts"☆21Jun 13, 2026Updated 3 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Offical repo for ECCV 2024: Depth-Aware Blind Image Decomposition for Real-World Weather Recovery☆13Mar 7, 2024Updated 2 years ago
- The official code of "Beyond Walking: A Large-Scale Image-Text Benchmark for Text-based Person Anomaly Search"☆38Jul 25, 2026Updated last month
- ☆16Jan 13, 2024Updated 2 years ago
- 🎨Official Repo for Every Painting Awakened: A Training-free Framework for Painting-to-Animation Generation☆57Apr 10, 2025Updated last year
- An offical repo for ECCV 2024 Towards Natural Language-Guided Drones: GeoText-1652 Benchmark with Spatial Relation Matching☆119Jul 7, 2026Updated 2 months ago
- UAVM @ ACM MM2023 Workshop on UAVs in Multimedia: Capturing the World from a New Perspective☆17Apr 30, 2025Updated last year
- ✨A curated list of papers on the uncertainty in multi-modal large language model (MLLM).☆59Apr 2, 2025Updated last year
- ☆11Mar 13, 2017Updated 9 years ago
- [ACM MM 2019] Code release for "Joint Adversarial Domain Adaptation" https://dl.acm.org/doi/10.1145/3343031.3351070☆19Sep 15, 2020Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Pytorch implementation of Each Part Matters: Local Patterns Facilitate Cross-view Geo-localization https://arxiv.org/abs/2008.11646☆99Jul 6, 2026Updated 2 months ago
- Dynamic Selective Network for RGB-D Salient Object Detection☆12Jan 22, 2025Updated last year
- World Model & VLA Survey - Interactive Research Page☆20May 26, 2026Updated 3 months ago
- [ACMMM 2020] Code release for "Simultaneous Semantic Alignment Network for Heterogenous Domain Adaptation" https://arxiv.org/abs/2008.01…☆32Apr 16, 2022Updated 4 years ago
- ☆13Nov 23, 2022Updated 3 years ago
- Attline: Visualize Attention at a Line in Diffusers☆28Apr 14, 2026Updated 4 months ago
- Official Implementation of Semantic Distribution-aware Contrastive Adaptation for Semantic Segmentation https://arxiv.org/abs/2105.05013☆38Apr 27, 2022Updated 4 years ago
- This is an implement of ACE: Adapting to Changing Environments for Semantic Segmentation☆15Jun 14, 2020Updated 6 years ago
- DeblurSR: Event-Based Motion Deblurring Under the Spiking Representation (AAAI 2024)☆30Nov 8, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- This is a project on visual spatial reasoning tasks-SIBench☆28Jan 12, 2026Updated 8 months ago
- Emoticon range output in UTF-8 using C☆15Jan 28, 2024Updated 2 years ago
- ☆15Aug 25, 2025Updated last year
- Progressive Text-to-3D Generation for Automatic 3D Prototyping (ACM TOMM)☆54Mar 14, 2026Updated 5 months ago
- [TPAMI 2021] Code release for "Generalized Domain Conditioned Adaptation Network" https://arxiv.org/abs/2103.12339☆45Jun 22, 2022Updated 4 years ago
- [ECCV2024] Reflective Instruction Tuning: Mitigating Hallucinations in Large Vision-Language Models☆20Jul 17, 2024Updated 2 years ago
- 2021语言与智能技术竞赛:机器阅读理解任务☆29Jun 13, 2021Updated 5 years ago
- The official implementation of Few-Cost Salient Object Detection with Adversarial-Paced Learning☆18May 3, 2024Updated 2 years ago
- Official Implementation of CoSMo: Content-Style Modulation for Image Retrieval with Text Feedback presented in CVPR 2021.☆65Sep 13, 2022Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A python script for downloading huggingface datasets and models.☆20Apr 10, 2025Updated last year
- We present **FOCI**, a benchmark for Fine-grained Object ClassIfication for large vision language models (LVLMs).☆19Jun 21, 2024Updated 2 years ago
- The official Pytorch implementation of paper "FedSoup: Improving Generalization and Personalization in Federated Learning via Selective M…☆18Apr 14, 2024Updated 2 years ago
- Codes and datasets of AAAI 2020 paper: Single camera training for person re-identification☆24Feb 4, 2020Updated 6 years ago
- ☆31Feb 8, 2023Updated 3 years ago
- ICLR‘24 Offical Implementation of Composed Image Retrieval with Text Feedback via Multi-grained Uncertainty Regularization☆74Jan 30, 2024Updated 2 years ago
- A simple and effective feature alignment method with proposed anchor loss for person re-identification☆15Aug 18, 2020Updated 6 years ago