Code implementation of our ICCV 2025 paper: On Large Multimodal Models as Open-World Image Classifiers
☆26Dec 4, 2025Updated 9 months ago
Alternatives and similar repositories for lmms-owc
Users that are interested in lmms-owc are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR '23 Highlight] Official repository for the paper "Quantum Multi-Model Fitting".☆12Mar 7, 2025Updated last year
- Official Implementation of MULTI-LANE (Multi Label class incremental learning via summarising pAtch tokeN Embeddings). Published in 3rd C…☆15Feb 20, 2025Updated last year
- Official implementation of "ConViS-Bench: Estimating Video Similarity Through Semantic Concepts", NeurIPS 2025☆27Nov 28, 2025Updated 9 months ago
- Official implementation of the CVPR '25 highlight paper "Compositional Caching for Training-free Open-vocabulary Attribute Detection"☆23Dec 23, 2024Updated last year
- [CVPR'25] Official implementation of the paper "Not Only Text: Exploring Compositionality of Visual Representations in Vision-Language Mo…☆18Nov 21, 2025Updated 9 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆60Jul 26, 2026Updated last month
- Code implementation of our NeurIPS 2023 paper: Vocabulary-free Image Classification☆107Feb 2, 2024Updated 2 years ago
- Official implementation of "Test-Time Zero-Shot Temporal Action Localization", CVPR 2024☆79Sep 11, 2024Updated last year
- [ECCV 2024] BUSCA: "Lost and Found: Overcoming Detector Failures in Online Multi-Object Tracking"☆45Dec 6, 2024Updated last year
- Official codebase for the paper "Training-Free Personalization via Retrieval and Reasoning on Fingerprints"☆25Nov 6, 2025Updated 10 months ago
- [NeurIPS '24] Frustratingly easy Test-Time Adaptation of VLMs!!☆63Mar 24, 2025Updated last year
- [CVPR 2024 Highlight] OpenBias: Open-set Bias Detection in Text-to-Image Generative Models☆26Feb 13, 2025Updated last year
- [WACV 26] Official code for the paper Safe Vision-Language Models via Unsafe Weights Manipulation☆16Mar 3, 2026Updated 6 months ago
- [CVPR-25🔥] Test-time Counterattacks (TTC) towards adversarial robustness of CLIP☆42Jun 4, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [CVPR Findings 2026] Large Multimodal Models as General In-Context Classifiers☆26Mar 1, 2026Updated 6 months ago
- Official Implementation of "CLEFT: Language-Image Contrastive Learning with Efficient Large Language Model and Prompt Fine-Tuning" on MIC…☆18Feb 12, 2025Updated last year
- Official implementation of "Harnessing Large Language Models for Training-free Video Anomaly Detection", CVPR 2024☆151Jul 15, 2024Updated 2 years ago
- Follow-Up Differential Descriptions: Language Models Resolve Ambiguities for Image Classification☆11Nov 15, 2023Updated 2 years ago
- [ICML 2024] "Envisioning Outlier Exposure by Large Language Models for Out-of-Distribution Detection"☆13Feb 15, 2025Updated last year
- Spatio-temporal object detector☆14Apr 22, 2021Updated 5 years ago
- [ICLR 2025] Adaptive prompt tailored pruning of T2I diffusion models.☆15Feb 1, 2025Updated last year
- Official repo for the TMLR paper "Discffusion: Discriminative Diffusion Models as Few-shot Vision and Language Learners"☆29Apr 27, 2024Updated 2 years ago
- Code for ICML 2023 paper "When and How Does Known Class Help Discover Unknown Ones? Provable Understandings Through Spectral Analysis"☆14Jun 24, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Code for BYOP [CVPR 2023]☆11Sep 25, 2023Updated 2 years ago
- [ICCV 2025] Superpowering Open-Vocabulary Object Detectors for X-ray Vision☆15Nov 3, 2025Updated 10 months ago
- [ICLR 2024] Test-Time RL with CLIP Feedback for Vision-Language Models.☆104Oct 20, 2025Updated 10 months ago
- Code for Fast as CHITA: Neural Network Pruning with Combinatorial Optimization☆14Aug 2, 2023Updated 3 years ago
- [CVPR 2025] Augmenting Multimodal LLMs with Self-Reflective Tokens for Knowledge-based Visual Question Answering☆57Jul 14, 2025Updated last year
- [WACV'23] Mixture Outlier Exposure for Out-of-Distribution Detection in Fine-grained Environments☆26Apr 12, 2023Updated 3 years ago
- [ICLR2023] NTK-SAP: Improving neural network pruning by aligning training dynamics☆20May 1, 2023Updated 3 years ago
- [NeurIPS 2025] Official Implementation for "Enhancing Vision-Language Model Reliability with Uncertainty-Guided Dropout Decoding"☆22Dec 8, 2024Updated last year
- ☆25Aug 1, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICLR 2026] "Inverse Virtual Try-On: Generating Multi-Category Product-Style Images from Clothed Individuals"☆50Mar 6, 2026Updated 6 months ago
- [ICLR'26] This repository is the implementation of "3D Aware Region Prompted Vision Language Model"☆31Feb 19, 2026Updated 6 months ago
- The Land-Diffuser is a novel application of the Denoising Diffusion Probabilistic Model (DDPM) in the realm of 3D Talking Head generation…☆13Dec 23, 2023Updated 2 years ago
- [ECCV 2024 Workshop🎈] The first agriculture benchmark to evaluate MM-LLMs.☆28Jan 1, 2025Updated last year
- [ECCV 2024] Teach CLIP to Develop a Number Sense for Ordinal Regression☆23Apr 1, 2025Updated last year
- (Pattern Recognition 2025) Towards Trustworthy Dataset Distillation☆14Dec 8, 2024Updated last year
- 2SSP: A Two-Stage Framework for Structured Pruning of LLMs☆22Aug 18, 2025Updated last year