☆60Nov 17, 2022Updated 3 years ago
Alternatives and similar repositories for image-retrieval-baseline
Users that are interested in image-retrieval-baseline are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆66Dec 15, 2023Updated 2 years ago
- ☆31Sep 7, 2022Updated 3 years ago
- ☆11Feb 23, 2023Updated 3 years ago
- code for TCL: Vision-Language Pre-Training with Triple Contrastive Learning, CVPR 2022☆270Oct 2, 2024Updated last year
- pytorch implementation of mvp: a multi-stage vision-language pre-training framework☆35Mar 1, 2023Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- KDD Cup 2020 Challenges for Modern E-Commerce Platform: Multimodalities Recall first place☆190Jul 22, 2020Updated 6 years ago
- Probing task; contextual embeddings -> textual definitions (EMNLP19)☆12Apr 22, 2021Updated 5 years ago
- ☆15Jul 24, 2017Updated 9 years ago
- Code for ALBEF: a new vision-language pre-training method☆1,755Sep 20, 2022Updated 3 years ago
- ☆170Nov 9, 2023Updated 2 years ago
- CCL2022 新闻脉络关系识别☆31Oct 4, 2022Updated 3 years ago
- [FGVC9-CVPR 2022] The second place solution for 2nd eBay eProduct Visual Search Challenge.☆26Aug 9, 2022Updated 4 years ago
- NTK scaled version of ALiBi position encoding in Transformer.☆69Aug 16, 2023Updated 2 years ago
- 人人都能看懂的轻量级解决方案☆15Jul 10, 2020Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Implementation of the Benchmark Approaches for Medical Instructional Video Classification (MedVidCL) and Medical Video Question Answering…☆31Jan 31, 2023Updated 3 years ago
- 2022京东 全球人工智能技术创新大赛 电商关键属性的图文匹配任务第1名方案☆35Jul 17, 2022Updated 4 years ago
- 2021腾讯广告算法大赛-赛道二-第五名方案☆20May 22, 2022Updated 4 years ago
- Code repository for Percival: a generalizable vision language foundation model for computed tomography☆15Jun 30, 2026Updated last month
- 对抗训练在NLP中的应用☆14Nov 22, 2021Updated 4 years ago
- METER: A Multimodal End-to-end TransformER Framework☆377Nov 16, 2022Updated 3 years ago
- ☆55May 14, 2020Updated 6 years ago
- TAP: Text-Aware Pre-training for Text-VQA and Text-Caption, CVPR 2021 (Oral)☆72May 22, 2023Updated 3 years ago
- Interactive visualization of Manifold-Constrained Hyper-Connections (mHC) for stable deep network training☆28Jul 4, 2026Updated last month
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Source code and data used in the papers ViQuAE (Lerner et al., SIGIR'22), Multimodal ICT (Lerner et al., ECIR'23) and Cross-modal Retriev…☆39Dec 19, 2024Updated last year
- An interactive storybook built with the help of ChatGPT and Stable Diffusion.☆13Jun 28, 2023Updated 3 years ago
- Cross-lingual image captioning☆93May 9, 2022Updated 4 years ago
- ☆13Sep 15, 2022Updated 3 years ago
- ☆62Oct 25, 2022Updated 3 years ago
- ☆27Dec 3, 2021Updated 4 years ago
- Medical ML Benchmark☆11May 16, 2023Updated 3 years ago
- Official repository of OFA (ICML 2022). Paper: OFA: Unifying Architectures, Tasks, and Modalities Through a Simple Sequence-to-Sequence L…☆2,558Apr 24, 2024Updated 2 years ago
- Joint Image and textual feature Fashion Style search☆12Apr 7, 2020Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Touchstone: Evaluating Vision-Language Models by Language Models☆84Jan 18, 2024Updated 2 years ago
- An official implementation for " UniVL: A Unified Video and Language Pre-Training Model for Multimodal Understanding and Generation"☆365Jul 25, 2024Updated 2 years ago
- Bling's Object detection tool☆55Jan 9, 2023Updated 3 years ago
- Codes for the EMNLP'2020 paper "Predicting Clinical Trial Results by Implicit Evidence Integration".☆14Jan 13, 2021Updated 5 years ago
- Chinese version of CLIP which achieves Chinese cross-modal retrieval and representation generation.☆167Nov 3, 2022Updated 3 years ago
- ☆73Mar 2, 2022Updated 4 years ago
- ☆15Nov 8, 2023Updated 2 years ago