[NeurIPS 2024 Spotlight ⭐️ & TPAMI 2025] Parameter-Inverted Image Pyramid Networks (PIIP)
☆113Aug 5, 2025Updated last year
Alternatives and similar repositories for PIIP
Users that are interested in PIIP are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [IGARSS 2025 Oral] A Simple Aerial Detection Baseline of Multimodal Language Models.☆92Feb 12, 2026Updated 5 months ago
- Learning 1D Causal Visual Representation with De-focus Attention Networks☆35Jun 7, 2024Updated 2 years ago
- [Remote Sensing 2026] Co-Training Vision Language Models for Remote Sensing Multi-task Learning☆38Jul 8, 2026Updated last month
- [TPAMI] Oriented object detection on STAR dataset.☆88Feb 3, 2025Updated last year
- The official implementation of ADDP (ICLR 2024)☆12Mar 27, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- GeoGround: A Unified Large Vision-Language Model for Remote Sensing Visual Grounding☆93May 10, 2025Updated last year
- ☆31Jun 29, 2022Updated 4 years ago
- ⚽️🤖 Benchmarking LLMs and deep-research agents on real-world football prediction — from the tactical "who scores in minute 67" to the st…☆22Jul 22, 2026Updated 2 weeks ago
- [IJCV] PointOBB-v3: Expanding Performance Boundaries of Single Point-Supervised Oriented Object Detection☆42Sep 25, 2025Updated 10 months ago
- [CVPR 2023]Implementation of Siamese Image Modeling for Self-Supervised Vision Representation Learning☆41Jun 6, 2024Updated 2 years ago
- [CVPR'25] Official repo of "Point2RBox-v2:Rethinking Point-supervised Oriented Object Detection with Spatial Layout Among Instances"☆43Aug 3, 2026Updated last week
- (NeurIPS 2024) Official repository of paper "Frozen-DETR: Enhancing DETR with Image Understanding from Frozen Foundation Models"☆34Mar 22, 2025Updated last year
- Adapter-X: A Novel General Parameter-Efficient Fine-Tuning Framework for Vision☆11Jul 22, 2024Updated 2 years ago
- [ECCV24] MOD-UV: Learning Mobile Object Detectors from Unlabeled Videos☆11Oct 7, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆21Jul 3, 2025Updated last year
- [ICLR'25] Official repo of "PointOBB-v2: Towards Simpler, Faster, and Stronger Single Point Supervised Oriented Object Detection"☆38Mar 27, 2025Updated last year
- [ICCV 2025] EA-ViT: Efficient Adaptation for Elastic Vision Transformer☆27Jul 28, 2025Updated last year
- ☆16Mar 26, 2025Updated last year
- [ICLR 2025 Spotlight] OmniCorpus: A Unified Multimodal Corpus of 10 Billion-Level Images Interleaved with Text☆426May 5, 2025Updated last year
- [ICLR'23] PyTorch Implementation for H2RBox☆105Feb 14, 2023Updated 3 years ago
- code for the paper Offline Prioritized Experience Replay☆12Jun 13, 2023Updated 3 years ago
- RISE-Video: Can Video Generators Decode Implicit World Rules?☆28Mar 26, 2026Updated 4 months ago
- [ECCV 2024] Official implementation of "LaMI-DETR: Open-Vocabulary Detection with Language Model Instruction"☆90Dec 23, 2025Updated 7 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Aerial Detection Toolbox☆11Jan 18, 2023Updated 3 years ago
- [ICLR 2026] SpaCE-10: A Comprehensive Benchmark for Multimodal Large Language Models in Compositional Spatial Intelligence☆20Jan 26, 2026Updated 6 months ago
- ☆27Oct 15, 2024Updated last year
- [AAAI'25 Oral] NightReID: A Large-Scale Nighttime Person Re-Identification Benchmark☆11Jun 10, 2025Updated last year
- Synthesizing Efficient Data with Diffusion Models for Person Re-Identification Pre-Training☆11Jan 23, 2024Updated 2 years ago
- ☆23Nov 29, 2024Updated last year
- Stepping VLMs onto the Court: Benchmarking Spatial Intelligence in Sports☆71Mar 15, 2026Updated 4 months ago
- INF-LLaVA: Dual-perspective Perception for High-Resolution Multimodal Large Language Model☆42Aug 4, 2024Updated 2 years ago
- ☆24Jan 22, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Official Implementation of "FP-DETR: Detection Transformer Advanced by Fully Pre-training"☆62Mar 30, 2022Updated 4 years ago
- ☆50Nov 10, 2023Updated 2 years ago
- ☆128Jul 29, 2024Updated 2 years ago
- ☆31Aug 3, 2023Updated 3 years ago
- A Survey on Vision-Language Geo-Foundation Models (VLGFMs)☆180May 24, 2025Updated last year
- [ICLR2025] γ -MOD: Mixture-of-Depth Adaptation for Multimodal Large Language Models☆45Oct 28, 2025Updated 9 months ago
- [ICLR 2025 Spotlight] Vision-RWKV: Efficient and Scalable Visual Perception with RWKV-Like Architectures☆556Feb 18, 2025Updated last year