OVAD: Open-vocabulary Attribute Detection code
☆30Aug 28, 2023Updated 3 years ago
Alternatives and similar repositories for ovad-benchmark-code
Users that are interested in ovad-benchmark-code are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- OvarNet official implement of the paper "OvarNet: Towards Open-vocabulary Object Attribute Recognition"☆104Apr 7, 2023Updated 3 years ago
- This repository provides data for the VAW dataset as described in the CVPR 2021 paper titled "Learning to Predict Visual Attributes in th…☆72Jul 22, 2022Updated 4 years ago
- Repository for the paper: Teaching Structured Vision & Language Concepts to Vision & Language Models☆47Sep 25, 2023Updated 2 years ago
- Code Implementation of "Unsupervised Recognition of Unknown Objects for Open-World Object Detection"☆33Oct 13, 2023Updated 2 years ago
- [ECCV 2024] ControlCap: Controllable Region-level Captioning☆81Oct 25, 2024Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code use to create COCO Attributes dataset and experiments in the associate ECCV 2016 paper.☆49Dec 26, 2022Updated 3 years ago
- ☆25Aug 1, 2023Updated 3 years ago
- EMNLP2023 - InfoSeek: A New VQA Benchmark focus on Visual Info-Seeking Questions☆27May 30, 2024Updated 2 years ago
- Official implementation and dataset for the NAACL 2024 paper "ComCLIP: Training-Free Compositional Image and Text Matching"☆37Aug 18, 2024Updated 2 years ago
- [AAAI 2026] Relation-R1: Progressively Cognitive Chain-of-Thought Guided Reinforcement Learning for Unified Relation Comprehension☆20Mar 6, 2026Updated 5 months ago
- [CVPR 2026] LongVideo-R1: Smart Navigation for Low-cost Long Video Understanding☆52Jul 7, 2026Updated last month
- Learning from Noisy Anchors for One-stage Object Detection☆27Apr 14, 2021Updated 5 years ago
- Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency☆62Jun 6, 2025Updated last year
- This repository contains the implementation and the building blocks of GlideNet and Informed Convolution. This work is published at CVPR …☆29Mar 7, 2022Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Find What You Want: Learning Demand-conditioned Object Attribute Space for Demand-driven Navigation☆63Jan 15, 2025Updated last year
- A pytorch Implementation of Open Vocabulary Object Detection with Pseudo Bounding-Box Labels☆65Jun 25, 2026Updated 2 months ago
- [ICML 2024] Fine-Grained Classes and How to Find Them☆14Jun 21, 2024Updated 2 years ago
- [CVPR23 Highlight] CREPE: Can Vision-Language Foundation Models Reason Compositionally?☆36Apr 27, 2023Updated 3 years ago
- [NeurIPS 2023] Zero-shot Visual Relation Detection via Composite Visual Cues from Large Language Models☆23Oct 21, 2025Updated 10 months ago
- Repository for SF2SE3: Clustering Scene Flow into SE(3)-Motions via Proposal and Selection☆12Jul 26, 2024Updated 2 years ago
- ☆10Aug 22, 2023Updated 3 years ago
- [NeurIPS2023] Official implementation of the paper "Large Language Models are Visual Reasoning Coordinators"☆106Nov 9, 2023Updated 2 years ago
- EagleVision: Object-level Attribute Multimodal LLM for Remote Sensing☆26May 29, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Out-of-Distribution detection for segmentation data, using deep nearest neighbors. Presented at the ICCV23 workshop on Uncertainty Quanti…☆15Oct 9, 2023Updated 2 years ago
- ☆21Jun 6, 2024Updated 2 years ago
- ☆10Apr 7, 2025Updated last year
- Semantic Enhanced Attribute Learning☆34Nov 13, 2023Updated 2 years ago
- Object-Aware Distillation Pyramid for Open-Vocabulary Object Detection☆64Jan 6, 2026Updated 7 months ago
- ☆101Jun 23, 2025Updated last year
- [CVPR 2024] Tune-An-Ellipse: CLIP Has Potential to Find What You Want☆14Jan 5, 2025Updated last year
- [NeurIPS2023] Official implementation and model release of the paper "What Makes Good Examples for Visual In-Context Learning?"☆182Mar 4, 2024Updated 2 years ago
- 📷 Neural Radiance Field research for synthesizing novel views☆18Mar 16, 2026Updated 5 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Code for the paper "Detecting Any Human-Object Interaction Relationship: Universal HOI Detector with Spatial Prompt Learning on Foundatio…☆28Nov 8, 2023Updated 2 years ago
- [CVPR 2023] Prompt, Generate, then Cache: Cascade of Foundation Models makes Strong Few-shot Learners☆45Jun 14, 2023Updated 3 years ago
- Official implementation of EgoThinker at NIPS 2025☆29Nov 25, 2025Updated 9 months ago
- [Pattern Recognition 25] CLIP Surgery for Better Explainability with Enhancement in Open-Vocabulary Tasks☆482Mar 1, 2025Updated last year
- [CVPR 2023] Official Pytorch code for PROB: Probabilistic Objectness for Open World Object Detection☆151Oct 29, 2024Updated last year
- [AAAI 2024]Weakly Supervised Multimodal Affordance Grounding for Egocentric Images☆13Nov 10, 2024Updated last year
- [NeurIPS 2024] Official PyTorch implementation of LoTLIP: Improving Language-Image Pre-training for Long Text Understanding☆49Jan 14, 2025Updated last year