The most impactful papers related to contrastive pretraining for multimodal models!
☆79Mar 5, 2024Updated 2 years ago
Alternatives and similar repositories for awesome-clip-papers
Users that are interested in awesome-clip-papers are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Remove exact and approximate duplicates from your dataset in FiftyOne!☆18Apr 4, 2024Updated 2 years ago
- Convert datasets from Hugging Face to FiftyOne for Visualization☆11Mar 15, 2024Updated 2 years ago
- Run optical character recognition with PyTesseract from the FiftyOne App!☆11Apr 5, 2024Updated 2 years ago
- Semantically Search Emojis From the Command Line!☆13Nov 26, 2023Updated 2 years ago
- Find the images in your dataset most similar to a query image from URL or drag-and-drop, with FiftyOne!☆15Nov 6, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [ICCV 2025] Official Repository for "Multi-Cache Enhanced Prototype Learning for Test-Time Generalization of Vision-Language Models"☆20Nov 10, 2025Updated 8 months ago
- Albumentations Data Augmentation Plugin for FiftyOne!☆15Aug 22, 2024Updated last year
- Testbed for multimodal retrieval augmented generation techniques with FiftyOne, LlamaIndex, and Milvus☆21Aug 9, 2024Updated last year
- Perform visual question answering on your images☆19May 8, 2024Updated 2 years ago
- [NeurIPS 2024] WATT: Weight Average Test-Time Adaptation of CLIP☆58Sep 26, 2024Updated last year
- [NeurIPS24] VisMin: Visual Minimal-Change Understanding☆19Mar 3, 2025Updated last year
- My journey during 10 weeks of building FiftyOne plugins☆22Nov 12, 2023Updated 2 years ago
- [ICLR 2025] "Noisy Test-Time Adaptation in Vision-Language Models"☆13Feb 22, 2025Updated last year
- [CVPR 2024 Highlight] Official repository of the paper "The devil is in the fine-grained details: Evaluating open-vocabulary object detec…☆68Apr 4, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A curated list of awesome prompt/adapter learning methods for vision-language models like CLIP.☆786Jul 17, 2026Updated last week
- CVPR2024: Dual Memory Networks: A Versatile Adaptation Approach for Vision-Language Models☆96Jul 4, 2024Updated 2 years ago
- WIKIGENBENCH: Exploring Full-length Wikipedia Generation under Real-World Scenario (COLING 2025)