Ranking-Consistent Language-Image Pretraining
☆15Oct 24, 2025Updated 10 months ago
Alternatives and similar repositories for RankCLIP
Users that are interested in RankCLIP are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for "CAFe: Unifying Representation and Generation with Contrastive-Autoregressive Finetuning"☆34Mar 26, 2025Updated last year
- RO-ViT CVPR 2023 "Region-Aware Pretraining for Open-Vocabulary Object Detection with Vision Transformers"☆17Aug 24, 2023Updated 3 years ago
- [NeurIPS 2025] Official Implementation for "Enhancing Vision-Language Model Reliability with Uncertainty-Guided Dropout Decoding"☆22Dec 8, 2024Updated last year
- Official code for the paper "Two Effects, One Trigger: On the Modality Gap, Object Bias, and Information Imbalance in Contrastive Vision-…☆26May 11, 2025Updated last year
- stable diffusion webui segment anything☆11Apr 11, 2023Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- A Large-Scale Chinese Image-Text Benchmark for Real-World Short Video Search Scenarios☆14Jan 24, 2024Updated 2 years ago
- ☆22Apr 27, 2024Updated 2 years ago
- Composed Video Retrieval☆62May 2, 2024Updated 2 years ago
- Official implementation for NeurIPS'23 paper "Geodesic Multi-Modal Mixup for Robust Fine-Tuning"☆35Sep 24, 2024Updated last year
- ☆117Jun 13, 2023Updated 3 years ago
- ☆30Jun 10, 2024Updated 2 years ago
- Repository to perform multi animal pose detection. In particular this code is used for bee pose estimation.☆10Jan 10, 2022Updated 4 years ago
- ☆24Aug 6, 2026Updated 3 weeks ago
- [ICML 2024] Official implementation for "HALC: Object Hallucination Reduction via Adaptive Focal-Contrast Decoding"☆116Dec 4, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆12Jun 1, 2023Updated 3 years ago
- Official Implementation of Attentive Mask CLIP (ICCV2023, https://arxiv.org/abs/2212.08653)☆38May 29, 2024Updated 2 years ago
- Data & Code for FEDD published @ MICCAI 23☆12Oct 11, 2023Updated 2 years ago
- A tool to identify and analyze open-source repositories affiliated with universities using GitHub metadata and contributor analysis.☆16Jul 28, 2026Updated last month
- This is the official repository for the publication https://arxiv.org/abs/2311.16682☆14Aug 10, 2024Updated 2 years ago
- ☆13May 10, 2025Updated last year
- Code for paper: Unified Text-to-Image Generation and Retrieval☆15Jul 19, 2026Updated last month
- Deep Learning - Visual Representation Learning by solving Jigsaw puzzles using Deep Reinforcement Learning☆10Dec 8, 2016Updated 9 years ago
- ☆11Feb 9, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Convolutional Fine-Grained Classification with Self-Supervised Target Relation Regularization (TIP 2022)☆12Sep 8, 2022Updated 3 years ago
- A simple Computer Vision Framework, mainly based on PyTorch. Including distributed training, logging and so on.☆12Dec 2, 2023Updated 2 years ago
- ☆15May 15, 2025Updated last year
- ☆81Oct 27, 2023Updated 2 years ago
- [Siggraph2025] The official code of the paper "ColorSurge: Bringing Vibrancy and Efficiency to Automatic Video Colorization via Dual-Bran…☆15Jul 26, 2025Updated last year
- [CBMI 2024 Best Paper] Official repository of the paper "Is CLIP the main roadblock for fine-grained open-world perception?".☆31May 12, 2025Updated last year
- [ICCV 2023] ViLLA: Fine-grained vision-language representation learning from real-world data☆45Oct 15, 2023Updated 2 years ago
- ☆45Aug 14, 2023Updated 3 years ago
- ☆17May 26, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [CVPR 2024] TeachCLIP for Text-to-Video Retrieval☆42May 7, 2025Updated last year
- This is Pytorch implementation of our paper "LF-ViT: Reducing Spatial Redundancy in Vision Transformer for Efficient Image Recognition".☆10Sep 23, 2024Updated last year
- Reasoning Guided Embeddings: Leveraging MLLM Reasoning for Improved Multimodal Retrieval☆16Nov 29, 2025Updated 9 months ago
- ☆10Aug 12, 2020Updated 6 years ago
- Using Vrep to simulate a six-legged robot to do motion planning & path planning☆10Jan 10, 2019Updated 7 years ago
- Semantic Graph Representation Learning for Handwritten Mathematical Expression Recognition (ICDAR 2023)☆14Aug 29, 2023Updated 3 years ago
- VideoEval: Comprehensive Benchmark Suite for Low-Cost Evaluation of Video Foundation Model☆15Jul 31, 2025Updated last year