(CVPR 2025 highlight✨) Official repository of paper "LLMDet: Learning Strong Open-Vocabulary Object Detectors under the Supervision of Large Language Models"
☆618Feb 4, 2026Updated 7 months ago
Alternatives and similar repositories for LLMDet
Users that are interested in LLMDet are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- (CVPR 2026) Official repository of paper "WeDetect: Fast Open-Vocabulary Object Detection as Retrieval"☆270Jun 7, 2026Updated 3 months ago
- [CVPR2026] Detect Anything via Next Point Prediction☆1,592Feb 22, 2026Updated 7 months ago
- [ECCV 2024] Official implementation of "LaMI-DETR: Open-Vocabulary Detection with Language Model Instruction"☆89Dec 23, 2025Updated 8 months ago
- (ICML 2026) Official repository of paper "ObjEmbed: Towards Universal Multimodal Object Embeddings"☆59May 18, 2026Updated 4 months ago
- [ICCV 2025] Official implementation of the paper: "Dynamic-DINO: Fine-Grained Mixture of Experts Tuning for Real-time Open-Vocabulary Obj…☆81Jul 29, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- YOLOE: Real-Time Seeing Anything [ICCV 2025]☆2,289Jun 26, 2025Updated last year
- YOLO-UniOW: Efficient Universal Open-World Object Detection☆195Jan 17, 2025Updated last year
- Make Large Multimodal Models excel in object detection, ICCV 2025☆65Aug 1, 2025Updated last year
- [CVPR 2024] Real-Time Open-Vocabulary Object Detection☆6,569Feb 26, 2025Updated last year
- OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion☆414Mar 12, 2025Updated last year
- [CVPR 2025] DeCLIP: Decoupled Learning for Open-Vocabulary Dense Perception☆228Jan 10, 2026Updated 8 months ago
- (NeurIPS 2024) Official repository of paper "Frozen-DETR: Enhancing DETR with Image Understanding from Frozen Foundation Models"☆34Mar 22, 2025Updated last year
- [CVPR 2025] Mr. DETR: Instructive Multi-Route Training for Detection Transformers☆175Sep 6, 2025Updated last year
- A curated list of papers and resources related to Described Object Detection, Open-Vocabulary/Open-World Object Detection and Referring E…☆360Nov 6, 2025Updated 10 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [CVPR 2025] DEIM: DETR with Improved Matching for Fast Convergence☆1,624Mar 24, 2026Updated 5 months ago
- DINO-X: The World's Top-Performing Vision Model for Open-World Object Detection and Understanding☆1,417Jul 23, 2025Updated last year
- [CVPR2024] Generative Region-Language Pretraining for Open-Ended Object Detection☆196Mar 29, 2025Updated last year
- [AAAI'25] Official Code for “Locate Anything on Earth: Advancing Open-Vocabulary Object Detection for Remote Sensing Community"☆295Jun 6, 2026Updated 3 months ago
- (TMM 2025) Official repository of paper "A Hierarchical Semantic Distillation Framework for Open-Vocabulary Object Detection"☆27Mar 14, 2025Updated last year
- Reference PyTorch implementation and models for DINOv3☆11,425Jul 15, 2026Updated 2 months ago
- ☆55Dec 23, 2024Updated last year
- This is the third party implementation of the paper Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detectio…☆844Jul 27, 2025Updated last year
- (TPAMI 2024) A Survey on Open Vocabulary Learning☆1,006May 12, 2026Updated 4 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [CVPR 2024] Official RT-DETR (RTDETR paddle pytorch), Real-Time DEtection TRansformer, DETRs Beat YOLOs on Real-time Object Detection. 🔥…☆5,548Sep 7, 2026Updated 2 weeks ago
- A curated list of papers, datasets and resources pertaining to open vocabulary object detection.☆424May 13, 2025Updated last year
- [ECCV 2024] Official implementation of the paper "Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection"☆10,608Aug 12, 2024Updated 2 years ago
- [DEIMv2] Real Time Object Detection Meets DINOv3☆2,040Aug 24, 2026Updated 3 weeks ago
- 📚 2025 Scene Graph ArXiv Paper List — Updated Daily☆16Mar 18, 2026Updated 6 months ago
- Offical implementation of "Re-Aligning Language to Visual Objects with an Agentic Workflow"☆34Apr 20, 2025Updated last year
- A benchmark for cross-domain few-shot object detection (ECCV24 paper: Cross-Domain Few-Shot Object Detection via Enhanced Open-Set Object…☆218Dec 10, 2025Updated 9 months ago
- [ECCV2024 Oral] Official implementation of the paper "Relation DETR: Exploring Explicit Position Relation Prior for Object Detection"☆265Nov 24, 2024Updated last year
- Official repo of Griffon series including v1(ECCV 2024), v2(ICCV 2025), G, and R, and also the RL tool Vision-R1(CVPR 2026).☆252Apr 17, 2026Updated 5 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [CVPR 2026 Highlight 🔥] PET-DINO: Unifying Visual Cues into Grounding DINO with Prompt-Enriched Training☆56May 6, 2026Updated 4 months ago
- [AAAI2025] Code Release of OV-DQUO: Open-Vocabulary DETR with Denoising Text Query Training and Open-World Unknown Objects Supervision☆138Dec 15, 2024Updated last year
- [ICCV2025] Harnessing CLIP, DINO and SAM for Open Vocabulary Segmentation☆129Nov 22, 2025Updated 10 months ago
- [ICLR 2026] Official implementation of "Patch-as-Decodable-Token: Towards Unified Multi-Modal Vision Tasks in MLLMs"☆165Aug 3, 2026Updated last month
- Solve Visual Understanding with Reinforced VLMs☆6,028Jul 7, 2026Updated 2 months ago
- Grounded SAM 2: Ground and Track Anything in Videos with Grounding DINO, Florence-2 and SAM 2☆3,744Nov 11, 2025Updated 10 months ago
- [CVPR2024 Highlight]GLEE: General Object Foundation Model for Images and Videos at Scale☆1,171Oct 21, 2024Updated last year