Code from the paper "Roboflow100-VL: A Multi-Domain Object Detection Benchmark for Vision-Language Models"
☆141Aug 14, 2026Updated last month
Alternatives and similar repositories for rf100-vl
Users that are interested in rf100-vl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆45Updated this week
- RF-DETR is a real-time object detection and segmentation model architecture developed by Roboflow, SOTA on COCO, designed for fine-tuning…☆9,551Updated this week
- Generalist YOLO: Towards Real-Time End-to-End Multi-Task Visual Language Models☆97Aug 25, 2026Updated 3 weeks ago
- Codes for ICML 2023 Learning Dynamic Query Combinations for Transformer-based Object Detection and Segmentation☆38Sep 12, 2023Updated 3 years ago
- Awesome paper for multi-modal llm with grounding ability☆21Oct 11, 2025Updated 11 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Description and applications of OpenAI's paper about DALL-E (2021) and implementation of other (CLIP-guided) zero-shot text-to-image gene…☆33Aug 11, 2022Updated 4 years ago
- BYTETrack, BoostTrack, BoostTrack+, and BoostTrack++ implemented in pure Rust☆17Sep 1, 2026Updated 3 weeks ago
- Turn any computer or edge device into a command center for your computer vision projects.☆2,456Updated this week
- Use Segment Anything 3 to label data for use with Autodistill.☆52Nov 25, 2025Updated 9 months ago
- [DEIMv2] Real Time Object Detection Meets DINOv3☆2,040Aug 24, 2026Updated 3 weeks ago
- [Under preparation] Code repo for "Open-Vocabulary DETR with Conditional Matching" (ECCV 2022)☆241Aug 3, 2022Updated 4 years ago
- YOLOE: Real-Time Seeing Anything [ICCV 2025]☆2,289Jun 26, 2025Updated last year
- Official code of the paper "VideoMolmo: Spatio-Temporal Grounding meets Pointing"☆57Jul 5, 2025Updated last year
- [ICCV 2025] Official implementation of the paper: "Dynamic-DINO: Fine-Grained Mixture of Experts Tuning for Real-time Open-Vocabulary Obj…☆81Jul 29, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- [TMLR 26] EdgeCrafter: Compact ViTs for Edge Dense Prediction via Task-Specialized Distillation☆343Aug 24, 2026Updated 3 weeks ago
- (CVPR 2025 highlight✨) Official repository of paper "LLMDet: Learning Strong Open-Vocabulary Object Detectors under the Supervision of La…☆618Feb 4, 2026Updated 7 months ago
- YOLO-UniOW: Efficient Universal Open-World Object Detection☆195Jan 17, 2025Updated last year
- [CVPR 2026] A Closer Look at Cross-Domain Few-Shot Object Detection: Fine-Tuning Matters and Parallel Decoder Helps☆56Aug 10, 2026Updated last month
- Unified, high-performance computer vision tracking library☆33Updated this week
- State-of-the-art Image & Video CLIP, Multimodal Large Language Models, and More!☆2,370Apr 13, 2026Updated 5 months ago
- [AAAI 2026] SimROD: A Simple Baseline for Raw Object Detection with Global and Local Enhancements☆31Nov 8, 2025Updated 10 months ago
- [CVPR 2025 Highlight] Official code and models for Encoder-only Mask Transformer (EoMT).☆625Jul 22, 2026Updated 2 months ago
- Interface with the Roboflow API and Python package for running inference (receiving predictions) and customizing result images from your …☆43May 13, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [CVPR2026] Detect Anything via Next Point Prediction☆1,592Feb 22, 2026Updated 7 months ago
- [AAAI'25] Official Code for “Locate Anything on Earth: Advancing Open-Vocabulary Object Detection for Remote Sensing Community"☆295Jun 6, 2026Updated 3 months ago
- ☆32Nov 17, 2025Updated 10 months ago
- [WACV 2021] Selective Spatio-Temporal Aggregation based Pose Refinement System: Towards understanding human activities in real-world vide…☆14Nov 4, 2021Updated 4 years ago
- Make Large Multimodal Models excel in object detection, ICCV 2025☆65Aug 1, 2025Updated last year
- ☆19May 16, 2025Updated last year
- Modular CLI pipeline for fine‑tuning RF‑DETR object detection models on custom datasets.☆35Dec 3, 2025Updated 9 months ago
- New testing protocol for learning local patch descriptors on Brown Phototour dataset☆17Apr 14, 2025Updated last year
- Code to reproduce results in the paper "Learning a Model for Inferring a Spatial Road Lane Network Graph using Self-Supervision" (ITSC 20…☆19Jul 6, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- This repository contains the implementation for the paper "Revisiting Few Shot Object Detection with Vision-Language Models"☆93May 30, 2025Updated last year
- Awesome Papers at the 2026 3D Vision Conference in Vancouver, BC☆19Apr 21, 2026Updated 5 months ago
- Official repo for our ICML 23 paper: "Multi-Modal Classifiers for Open-Vocabulary Object Detection"☆95Jun 22, 2023Updated 3 years ago
- [CVPR 2025] Official PyTorch implementation of "EdgeTAM: On-Device Track Anything Model"☆980Jan 27, 2026Updated 7 months ago
- [CVPR 2025] DEIM: DETR with Improved Matching for Fast Convergence☆1,624Mar 24, 2026Updated 5 months ago
- ☆14Jan 26, 2025Updated last year
- Video Benchmark Suite: Rapid Evaluation of Video Foundation Models☆17Jan 10, 2025Updated last year