Code from the paper "Roboflow100-VL: A Multi-Domain Object Detection Benchmark for Vision-Language Models"
☆137Jul 8, 2026Updated 2 weeks ago
Alternatives and similar repositories for rf100-vl
Users that are interested in rf100-vl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆44Jun 11, 2026Updated last month
- RF-DETR is a real-time object detection and segmentation model architecture developed by Roboflow, SOTA on COCO, designed for fine-tuning…☆8,658Updated this week
- ☆12Nov 4, 2024Updated last year
- Codes for ICML 2023 Learning Dynamic Query Combinations for Transformer-based Object Detection and Segmentation☆38Sep 12, 2023Updated 2 years ago
- Awesome paper for multi-modal llm with grounding ability☆21Oct 11, 2025Updated 9 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Description and applications of OpenAI's paper about DALL-E (2021) and implementation of other (CLIP-guided) zero-shot text-to-image gene…☆33Aug 11, 2022Updated 3 years ago
- Turn any computer or edge device into a command center for your computer vision projects.☆2,390Updated this week
- [DEIMv2] Real Time Object Detection Meets DINOv3☆1,942Mar 24, 2026Updated 4 months ago
- [Under preparation] Code repo for "Open-Vocabulary DETR with Conditional Matching" (ECCV 2022)☆240Aug 3, 2022Updated 3 years ago
- YOLOE: Real-Time Seeing Anything [ICCV 2025]☆2,214Jun 26, 2025Updated last year
- An extension of RF-DETR that unlocks larger, more powerful detection models for when you need the highest possible accuracy.☆50Jul 9, 2026Updated 2 weeks ago
- Official code of the paper "VideoMolmo: Spatio-Temporal Grounding meets Pointing"☆56Jul 5, 2025Updated last year
- [ICCV 2025] Official implementation of the paper: "Dynamic-DINO: Fine-Grained Mixture of Experts Tuning for Real-time Open-Vocabulary Obj…☆79Jul 29, 2025Updated 11 months ago
- PyTorch implementation of "UNIT: Unifying Image and Text Recognition in One Vision Encoder", NeurlPS 2024.☆34Sep 26, 2024Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Code for the paper "Benchmarking Object Detectors with COCO: A New Path Forward."☆36Jul 13, 2024Updated 2 years ago
- (CVPR 2025 highlight✨) Official repository of paper "LLMDet: Learning Strong Open-Vocabulary Object Detectors under the Supervision of La…☆607Feb 4, 2026Updated 5 months ago
- [CVPR 2026] A Closer Look at Cross-Domain Few-Shot Object Detection: Fine-Tuning Matters and Parallel Decoder Helps☆47May 11, 2026Updated 2 months ago
- ☆15Oct 24, 2025Updated 8 months ago
- State-of-the-art Image & Video CLIP, Multimodal Large Language Models, and More!☆2,330Apr 13, 2026Updated 3 months ago
- [AAAI 2026] SimROD: A Simple Baseline for Raw Object Detection with Global and Local Enhancements☆31Nov 8, 2025Updated 8 months ago
- [CVPR 2025 Highlight] Official code and models for Encoder-only Mask Transformer (EoMT).☆615Updated this week
- [CVPR2026] Detect Anything via Next Point Prediction☆1,516Feb 22, 2026Updated 5 months ago
- [AAAI'25] Official Code for “Locate Anything on Earth: Advancing Open-Vocabulary Object Detection for Remote Sensing Community"☆277Jun 6, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- This repository is an official implementation of the paper "LW-DETR: A Transformer Replacement to YOLO for Real-Time Detection".☆506Feb 18, 2025Updated last year
- [WACV 2021] Selective Spatio-Temporal Aggregation based Pose Refinement System: Towards understanding human activities in real-world vide…☆13Nov 4, 2021Updated 4 years ago
- Pytorch implementation of "EdgeCrafter: Compact ViTs for Edge Dense Prediction via Task-Specialized Distillation"☆288Updated this week
- Make Large Multimodal Models excel in object detection, ICCV 2025☆65Aug 1, 2025Updated 11 months ago
- ☆16May 16, 2025Updated last year
- Modular CLI pipeline for fine‑tuning RF‑DETR object detection models on custom datasets.☆35Dec 3, 2025Updated 7 months ago
- Universal annotation converter☆17Updated this week
- This repository contains the implementation for the paper "Revisiting Few Shot Object Detection with Vision-Language Models"☆93May 30, 2025Updated last year
- Production-ready C++/TensorRT inference engine for RF-DETR. Object detection and instance segmentation with FP32/FP16/INT8 support. Optim…☆188Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code for replicating Roboflow 100 benchmark results and programmatically downloading benchmark datasets☆303Oct 26, 2024Updated last year
- [CVPR 2025] Official PyTorch implementation of "EdgeTAM: On-Device Track Anything Model"☆949Jan 27, 2026Updated 5 months ago
- [CVPR 2025] DEIM: DETR with Improved Matching for Fast Convergence☆1,589Mar 24, 2026Updated 4 months ago
- Video Benchmark Suite: Rapid Evaluation of Video Foundation Models☆17Jan 10, 2025Updated last year
- MMPD Dataset from ECCV'2024 "When Pedestrian Detection Meets Multi-Modal Learning: Generalist Model and Benchmark Dataset"☆22Jul 15, 2024Updated 2 years ago
- The repository provides code for running inference and finetuning with the Meta Segment Anything Model 3 (SAM 3), links for downloading t…☆11,063Jul 15, 2026Updated last week
- ☆19Nov 6, 2022Updated 3 years ago