[CVPR'25] Official implementation of the paper "Not Only Text: Exploring Compositionality of Visual Representations in Vision-Language Models".
☆18Nov 21, 2025Updated 9 months ago
Alternatives and similar repositories for vlm_image_compositionality
Users that are interested in vlm_image_compositionality are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR '23 Highlight] Official repository for the paper "Quantum Multi-Model Fitting".☆12Mar 7, 2025Updated last year
- [WACV 26] Official code for the paper Safe Vision-Language Models via Unsafe Weights Manipulation☆16Mar 3, 2026Updated 5 months ago
- Official implementation of "ConViS-Bench: Estimating Video Similarity Through Semantic Concepts", NeurIPS 2025☆27Nov 28, 2025Updated 9 months ago
- Official repo of the paper “AL-GTD: Deep Active Learning for Gaze Target Detection” (ACMMM2024)☆12Jul 23, 2026Updated last month
- [CVPR Findings 2026] Large Multimodal Models as General In-Context Classifiers☆25Mar 1, 2026Updated 5 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [CVPR'26 Highlight] MemCoach: Steering-based MLLM for Actionable Image Memorability Feedback☆44Jul 24, 2026Updated last month
- Code implementation of our ICCV 2025 paper: On Large Multimodal Models as Open-World Image Classifiers☆26Dec 4, 2025Updated 8 months ago
- ☆60Jul 26, 2026Updated last month
- Official repository of "ProactiveBench: Benchmarking Proactiveness in Multimodal Large Language Models" (ECCV 2026)☆31Jun 22, 2026Updated 2 months ago
- [CVPR 2024 Highlight] OpenBias: Open-set Bias Detection in Text-to-Image Generative Models☆26Feb 13, 2025Updated last year
- [ICPR 2024] Exemplar-free continual deepfake detector that leverages CLIP and domain-specific multi-modal prompts☆15Aug 1, 2024Updated 2 years ago
- [ICLR2023] NTK-SAP: Improving neural network pruning by aligning training dynamics☆20May 1, 2023Updated 3 years ago
- Official implementation of "Test-Time Zero-Shot Temporal Action Localization", CVPR 2024☆79Sep 11, 2024Updated last year
- [CVPR '25] Official implementation of the paper "Rethinking Few-Shot Adaptation of Vision-Language Models in Two Stages", CVPR 2025.☆33Mar 30, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [CVPR-25🔥] Test-time Counterattacks (TTC) towards adversarial robustness of CLIP☆41Jun 4, 2025Updated last year
- 2SSP: A Two-Stage Framework for Structured Pruning of LLMs☆22Aug 18, 2025Updated last year
- [ICLR 2026] "Inverse Virtual Try-On: Generating Multi-Category Product-Style Images from Clothed Individuals"☆49Mar 6, 2026Updated 5 months ago
- Code for "Are “Hierarchical” Visual Representations Hierarchical?" in NeurIPS Workshop for Symmetry and Geometry in Neural Representation…☆23Nov 8, 2023Updated 2 years ago
- [ICLR2023] Video Scene Graph Generation from Single-Frame Weak Supervision☆12Sep 17, 2023Updated 2 years ago
- PyTorch Implementation of the paper "Defining and Quantifying the Emergence of Sparse Concepts in DNNs" (CVPR 2023)☆12Dec 24, 2023Updated 2 years ago
- Reasoning Guided Embeddings: Leveraging MLLM Reasoning for Improved Multimodal Retrieval☆16Nov 29, 2025Updated 9 months ago
- [ICLR 2025] Adaptive prompt tailored pruning of T2I diffusion models.☆15Feb 1, 2025Updated last year
- ☆13Jun 11, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [FG 2026] Official implementation of the paper "NullFace: Training-Free Localized Face Anonymization"☆30Apr 28, 2026Updated 4 months ago
- An official implementation for MS-DETR in ACL'23☆17Jun 3, 2023Updated 3 years ago
- [ECCV 2024] BUSCA: "Lost and Found: Overcoming Detector Failures in Online Multi-Object Tracking"☆45Dec 6, 2024Updated last year
- [CVPR 23] Q: How to Specialize Large Vision-Language Models to Data-Scarce VQA Tasks? A: Self-Train on Unlabeled Images!☆17May 14, 2024Updated 2 years ago
- Generating Structured Pseudo Labels for Noise-resistant Zero-shot Video Sentence Localization☆16Jul 20, 2023Updated 3 years ago
- [ECCV2024] Reflective Instruction Tuning: Mitigating Hallucinations in Large Vision-Language Models☆20Jul 17, 2024Updated 2 years ago
- 武汉大学遥感院摄影测量学作业 Homework of Course Photogrammetry in Wuhan University: Space-Resection(Photogrammetry) 空间后方交会☆11Jan 11, 2022Updated 4 years ago
- Pytorch Implementation of ECCV'22 paper: Video Activity Localisation with Uncertainties in Temporal Boundary☆17Jul 17, 2022Updated 4 years ago
- (CVPR2026 Highlight) The Devil Is in Gradient Entanglement: Energy-Aware Gradient Coordinator for Robust Generalized Category Discovery (…☆33Aug 10, 2026Updated 2 weeks ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- [ICLR 2025] Data-Augmented Phrase-Level Alignment for Mitigating Object Hallucination☆21Jan 27, 2025Updated last year
- Official implementation of "Harnessing Large Language Models for Training-free Video Anomaly Detection", CVPR 2024☆151Jul 15, 2024Updated 2 years ago
- [CVPR 2025] DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval☆22Jun 23, 2025Updated last year
- Code for the paper "Manipulating Embeddings of Stable Diffusion Prompts".☆15Aug 8, 2024Updated 2 years ago
- Code for Fast as CHITA: Neural Network Pruning with Combinatorial Optimization☆14Aug 2, 2023Updated 3 years ago
- The Source Code for IF-VidCap @ICLR 2026☆18Oct 22, 2025Updated 10 months ago
- Source code of the paper Dual Learning with Dynamic Knowledge Distillation and Soft Alignment for Partially Relevant Video Retrieval☆19May 13, 2026Updated 3 months ago