Official implementation of the CVPR '25 highlight paper "Compositional Caching for Training-free Open-vocabulary Attribute Detection"
☆23Dec 23, 2024Updated last year
Alternatives and similar repositories for ComCa
Users that are interested in ComCa are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [WACV 26] Official code for the paper Safe Vision-Language Models via Unsafe Weights Manipulation☆16Mar 3, 2026Updated 5 months ago
- Official Implementation of MULTI-LANE (Multi Label class incremental learning via summarising pAtch tokeN Embeddings). Published in 3rd C…☆15Feb 20, 2025Updated last year
- [CVPR '23 Highlight] Official repository for the paper "Quantum Multi-Model Fitting".☆12Mar 7, 2025Updated last year
- Code implementation of our ICCV 2025 paper: On Large Multimodal Models as Open-World Image Classifiers☆26Dec 4, 2025Updated 8 months ago
- Official codebase for the paper "Training-Free Personalization via Retrieval and Reasoning on Fingerprints"☆25Nov 6, 2025Updated 9 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [CVPR'26 Highlight] MemCoach: Steering-based MLLM for Actionable Image Memorability Feedback☆43Jul 24, 2026Updated 3 weeks ago
- [CVPR'25] Official implementation of the paper "Not Only Text: Exploring Compositionality of Visual Representations in Vision-Language Mo…☆18Nov 21, 2025Updated 8 months ago
- [CVPR Findings 2026] Large Multimodal Models as General In-Context Classifiers☆25Mar 1, 2026Updated 5 months ago
- (CVPR2026 Highlight) The Devil Is in Gradient Entanglement: Energy-Aware Gradient Coordinator for Robust Generalized Category Discovery (…☆33Aug 10, 2026Updated last week
- [CVPR 2024 Highlight] OpenBias: Open-set Bias Detection in Text-to-Image Generative Models☆26Feb 13, 2025Updated last year
- [CVPR '25] Official implementation of the paper "Rethinking Few-Shot Adaptation of Vision-Language Models in Two Stages", CVPR 2025.☆33Mar 30, 2025Updated last year
- [ECCV 2024] BUSCA: "Lost and Found: Overcoming Detector Failures in Online Multi-Object Tracking"☆44Dec 6, 2024Updated last year
- [FG 2026] Official implementation of the paper "NullFace: Training-Free Localized Face Anonymization"☆29Apr 28, 2026Updated 3 months ago
- [TCSVT23] Official code for "SPT: Spatial Pyramid Transformer for Image Captioning".☆10Aug 14, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [ICPR 2024] Exemplar-free continual deepfake detector that leverages CLIP and domain-specific multi-modal prompts☆15Aug 1, 2024Updated 2 years ago
- [NeurIPS '24] Frustratingly easy Test-Time Adaptation of VLMs!!☆64Mar 24, 2025Updated last year
- Code for ICCV 2023 paper ✨ "StylerDALLE: Language-Guided Style Transfer Using a Vector-Quantized Tokenizer of a Large-Scale Generative Mo…☆18Jan 25, 2024Updated 2 years ago
- [ICLR 2025] Official repository of "Learning Clustering-based Prototypes for Compositional Zero-shot Learning"☆25Feb 3, 2026Updated 6 months ago
- Pytorch implementation of "Diversified in-domain synthesis with efficient fine-tuning for few-shot classification"☆17Mar 25, 2024Updated 2 years ago
- Loomis Painter: Reconstructing the painting process☆55Nov 24, 2025Updated 8 months ago
- ☆58Jul 26, 2026Updated 3 weeks ago
- ☆28Oct 31, 2024Updated last year
- [3DV 2026] Dense Motion Captioning☆36Jan 28, 2026Updated 6 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [CVPR-25🔥] Test-time Counterattacks (TTC) towards adversarial robustness of CLIP☆41Jun 4, 2025Updated last year
- Code for Fast as CHITA: Neural Network Pruning with Combinatorial Optimization☆14Aug 2, 2023Updated 3 years ago
- Official implementation of "What does CLIP know about a red circle? Visual Prompt Engineering for VLMs", ICCV 2023☆12Sep 21, 2023Updated 2 years ago
- Appunti del corso di Linguaggi formali e compilatori tenutosi l'A.A. 2018-2019.☆25Jan 28, 2020Updated 6 years ago
- (ICLR2026) Efficient Degradation-agnostic Image Restoration via Channel-Wise Functional Decomposition and Manifold Regularization☆35Jun 23, 2026Updated last month
- Official implementation of "Harnessing Large Language Models for Training-free Video Anomaly Detection", CVPR 2024☆151Jul 15, 2024Updated 2 years ago
- [AAAI'25]: Building a Multi-modal Spatiotemporal Expert for Zero-shot Action Recognition with CLIP☆23Aug 5, 2025Updated last year
- Official implementation of the WACV 2025 paper "3D Part Segmentation via Geometric Aggregation of 2D Visual Features"☆25Jun 8, 2025Updated last year
- ☆11Jul 11, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆14Jan 5, 2022Updated 4 years ago
- Official implementations of our LaZSL (ICCV'25)☆45Jul 13, 2025Updated last year
- 🔥🔥A Family of Multi-Sensor, Multi-Granularity Vision-Language Models for Earth Observation Understanding☆137Jun 15, 2026Updated 2 months ago
- Code for WACV 2024 paper ✨ "SpectralCLIP: Preventing Artifacts in Text-Guided Style Transfer from a Spectral Perspective".☆19Nov 4, 2023Updated 2 years ago
- Official implementation of "In-style: Bridging Text and Uncurated Videos with Style Transfer for Cross-modal Retrieval." ICCV 2023☆11Oct 5, 2023Updated 2 years ago
- The implementation of Learning Instance and Task-Aware Dynamic Kernels for Few Shot Learning☆13Apr 14, 2024Updated 2 years ago
- [NeurIPS 2024] Official Code for the Paper "Multimodal Task Vectors Enable Many-Shot Multimodal In-Context Learning"☆27Apr 8, 2025Updated last year