Implementation of the "Learn No to Say Yes Better" paper.
☆40Apr 3, 2026Updated 3 months ago
Alternatives and similar repositories for CoN-CLIP
Users that are interested in CoN-CLIP are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Project for SNARE benchmark☆11Jun 5, 2024Updated 2 years ago
- [CVPR23 Highlight] CREPE: Can Vision-Language Foundation Models Reason Compositionally?☆35Apr 27, 2023Updated 3 years ago
- Code repository for "Post-pre-training for Modality Alignment in Vision-Language Foundation Models" (CVPR2025)☆41Jul 25, 2025Updated 11 months ago
- Code and data release for the paper "Learning Fine-grained View-Invariant Representations from Unpaired Ego-Exo Videos via Temporal Align…☆19Apr 5, 2024Updated 2 years ago
- ☆44Apr 8, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Repository for the paper: Teaching Structured Vision & Language Concepts to Vision & Language Models☆47Sep 25, 2023Updated 2 years ago
- Latest Advances on Modality Priors in Multimodal Large Language Models☆30Dec 10, 2025Updated 7 months ago
- ☆13Aug 14, 2022Updated 3 years ago
- [ICML'24 Oral] "MagicLens: Self-Supervised Image Retrieval with Open-Ended Instructions"☆211Oct 28, 2024Updated last year
- Sets of Image Provenance cases, including node and edge information, generated automatically using Reddit Photoshop Battles☆13Jul 26, 2018Updated 7 years ago
- Repo for "Synergy of Sight and Semantics: Visual Intention Understanding with CLIP"☆12Mar 12, 2025Updated last year
- Code and data release for the paper "Learning Object State Changes in Videos: An Open-World Perspective" (CVPR 2024)☆37Sep 9, 2024Updated last year
- [NeurIPS 2024] Official PyTorch implementation of "Improving Compositional Reasoning of CLIP via Synthetic Vision-Language Negatives"☆48Dec 1, 2024Updated last year
- Smooth Variational Graph Embeddings for Efficient Neural Architecture Search☆14Feb 2, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- COLA: Evaluate how well your vision-language model can Compose Objects Localized with Attributes!☆25May 14, 2026Updated 2 months ago
- ☆28Jul 18, 2025Updated last year
- VisualGPTScore for visio-linguistic reasoning☆27Oct 7, 2023Updated 2 years ago
- Personal Claude Code plugin marketplace☆16Jul 4, 2026Updated 2 weeks ago
- Towards a Unified View on Visual Parameter-Efficient Transfer Learning☆26Oct 13, 2022Updated 3 years ago
- NegCLIP.☆41Feb 6, 2023Updated 3 years ago
- The official repository for CosPGD: a unified white-box adversarial attack for pixel-wise prediction tasks.☆15May 8, 2025Updated last year
- ☆13Apr 12, 2026Updated 3 months ago
- Smooth Variational Graph Embeddings for Efficient Neural Architecture Search☆15Apr 8, 2024Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- [ICLR 2024] Official repository for "Vision-by-Language for Training-Free Compositional Image Retrieval"☆89Jul 4, 2024Updated 2 years ago
- CLIP-MoE: Mixture of Experts for CLIP☆58Oct 10, 2024Updated last year
- Open-source strong baseline for domain generlization re-ID. We will udpate the strong baseline and CFD method~☆10Nov 30, 2021Updated 4 years ago
- Official repository for the ICCV 2023 paper: "Waffling around for Performance: Visual Classification with Random Words and Broad Concepts…☆61Jul 8, 2023Updated 3 years ago
- This is the official implementation of our BMVC 2022 paper "SP-ViT: Learning 2D Spatial Priors for Vision Transformers"☆14Mar 27, 2023Updated 3 years ago
- [CVPR 2025 Highlight] Official Pytorch codebase for paper: "Assessing and Learning Alignment of Unimodal Vision and Language Models"☆60Aug 15, 2025Updated 11 months ago
- This repo contains the data used in "Towards Understanding Climate Change Perceptions: A Social Media Dataset"☆15Apr 13, 2026Updated 3 months ago
- GitHub repository of the ICLR 2023 paper "Neural Architecture Design and Robustness: A Dataset"☆16Jan 25, 2023Updated 3 years ago
- Code for the paper "AMEGO: Active Memory from long EGOcentric videos" published at ECCV 2024☆45Dec 7, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Code repository for paper: "General surgery vision transformer: A video pre-trained foundation model for general surgery"☆51Apr 19, 2024Updated 2 years ago
- Official implementation of "In-style: Bridging Text and Uncurated Videos with Style Transfer for Cross-modal Retrieval." ICCV 2023☆11Oct 5, 2023Updated 2 years ago
- Experiments and data for the paper "When and why vision-language models behave like bags-of-words, and what to do about it?" Oral @ ICLR …☆294Jun 7, 2023Updated 3 years ago
- CLiC: Concept Learning in Context☆10Jan 24, 2025Updated last year
- [CVPR 2026 Findings] PDF-GS: Progressive Distractor Filtering for Robust 3D Gaussian Splatting☆15Jun 6, 2026Updated last month
- ☆14Dec 12, 2023Updated 2 years ago
- AlignCLIP: Improving Cross-Modal Alignment in CLIP (ICLR 2025)☆67Mar 1, 2025Updated last year