ContextBLIP : Doubly Contextual Alignment for Contrastive Image Retrieval from Linguistically Complex Descriptions [ACL 2024]
☆11May 17, 2024Updated 2 years ago
Alternatives and similar repositories for ContextBLIP
Users that are interested in ContextBLIP are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MetaLadder: Ascending Mathematical Solution Quality via Analogical-Problem Reasoning Transfer (EMNLP 2025)☆12Apr 18, 2025Updated last year
- Software used to automatically calibrate the extrinsic parameters of several sensors of different modalities☆11Jun 12, 2015Updated 11 years ago
- [ICML 2026] What Does Vision Tool-Use Reinforcement Learning Really Learn? Disentangling Tool-Induced and Intrinsic Effects for Crop-and-…☆23May 15, 2026Updated 3 months ago
- [NeurIPS 2023] Bootstrapping Vision-Language Learning with Decoupled Language Pre-training☆26Dec 5, 2023Updated 2 years ago
- MaXM is a suite of test-only benchmarks for multilingual visual question answering in 7 languages: English (en), French (fr), Hindi (hi),…☆13Jan 16, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- 北邮福建群志☆12Nov 17, 2022Updated 3 years ago
- Mahalanobis Distance-based Multi-view Optimal Transport for Multi-view Crowd Localization, ECCV 2024☆15Nov 20, 2024Updated last year
- ☆12Jun 21, 2022Updated 4 years ago
- Use Siamese Network to implement fingerprint verification task.☆11Oct 21, 2021Updated 4 years ago
- ☆12Jan 16, 2024Updated 2 years ago
- The official pytorch implementation of Exploring the Interactive Guidance for Unified and Effective Image Matting [TOMM 2025]☆25Nov 24, 2025Updated 9 months ago
- [EMNLP 2024] Preserving Multi-Modal Capabilities of Pre-trained VLMs for Improving Vision-Linguistic Compositionality☆24Oct 8, 2024Updated last year
- Trace origins, shared sources, and contamination risk☆27May 27, 2026Updated 3 months ago
- Code for paper: Unified Text-to-Image Generation and Retrieval☆15Jul 19, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Vehicle registration plate recognition using convolutional neural networks☆11Nov 30, 2022Updated 3 years ago
- ☆15Jul 9, 2024Updated 2 years ago
- PyTorch implementation of LIMoE☆52Apr 1, 2024Updated 2 years ago
- [NeurIPS 2025] Reasoning MLLM, Share-GRPO, advantage vanishing, sparse reward☆38Sep 19, 2025Updated 11 months ago
- KAIST medical VL research group☆20Dec 20, 2024Updated last year
- PyTorch implementation of StackGAN paper using BERT embeddings☆12Feb 6, 2022Updated 4 years ago
- This is a repository for the paper "Transformer-based Missing Well Log Prediction".☆18Nov 6, 2023Updated 2 years ago
- Language-Guided Face Animation by Recurrent StyleGAN-based Generator☆20Apr 23, 2023Updated 3 years ago
- This project aims to collect and collate various datasets for multimodal large model training, including but not limited to pre-training …☆78May 7, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- 基于文字密度的新闻正文提取模块,兼容python2和python3,传入新闻网址或者网页源码即可返回标题,发布时间和正文内容。☆14Jun 10, 2018Updated 8 years ago
- ☆12Feb 2, 2023Updated 3 years ago
- ☆20Mar 5, 2025Updated last year
- MathFusion: Enhancing Mathematical Problem-solving of LLM through Instruction Fusion (ACL 2025)☆37Jul 16, 2025Updated last year
- [CVPR 2025] PyTorch implementation of paper "FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training"☆33Jul 8, 2025Updated last year
- A toy example for RNN in Python☆19Feb 29, 2016Updated 10 years ago
- [ICLR 2025] Official code repository for "TULIP: Token-length Upgraded CLIP"☆32Jan 26, 2026Updated 7 months ago
- Cognitive Science 2 exam project☆11Jun 3, 2018Updated 8 years ago
- Code for paper: "Region Proposals for Saliency Map Refinement for Weakly-supervised Disease Localisation and Classification"☆14Jun 29, 2021Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Implementation of BadCLIP https://arxiv.org/pdf/2311.16194.pdf☆25Mar 23, 2024Updated 2 years ago
- Repo for ICCV 2021 paper: Beyond Question-Based Biases: Assessing Multimodal Shortcut Learning in Visual Question Answering☆29Jul 1, 2024Updated 2 years ago
- In this work, we implement different cross-modal learning schemes such as Siamese Network, Correlational Network and Deep Cross-Modal Pro…☆11Aug 23, 2021Updated 5 years ago
- This is the repo for the paper Multi-Agent Collaborative Data Selection for Efficient LLM Pretraining.☆49Aug 22, 2025Updated last year
- ☆24Apr 16, 2025Updated last year
- Trying to classify the 20BN-JESTER hand gesture data set using a few architectures.☆17May 8, 2018Updated 8 years ago
- ☆12Sep 23, 2022Updated 3 years ago