CLIPxGPT Captioner is Image Captioning Model based on OpenAI's CLIP and GPT-2.
☆117Feb 17, 2025Updated last year
Alternatives and similar repositories for clip-gpt-captioning
Users that are interested in clip-gpt-captioning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV 2023] Going Beyond Nouns With Vision & Language Models Using Synthetic Data☆13Sep 30, 2023Updated 2 years ago
- An up-to-date & curated list of awesome layout to image papers, methods & resources.☆13Jun 28, 2024Updated 2 years ago
- Retrieval-augmented Image Captioning☆13Feb 16, 2023Updated 3 years ago
- Simple image captioning model☆1,423Jun 9, 2024Updated 2 years ago
- This repository contains the code and datasets for our ICCV-W paper 'Enhancing CLIP with GPT-4: Harnessing Visual Descriptions as Prompts…☆30Feb 21, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Timecode and chapter generator for YouTube videos based on heuristic scene detection and CLIPxGPT Captioner☆13Mar 13, 2025Updated last year
- ☆11May 5, 2024Updated 2 years ago
- finetune script for SDXL adapted from waifu-diffusion trainer☆11Aug 21, 2023Updated 3 years ago
- GRIT: Faster and Better Image-captioning Transformer (ECCV 2022)☆199May 9, 2023Updated 3 years ago
- Create your own DALL-E application in Python with Streamlit.☆12Mar 9, 2023Updated 3 years ago
- CapDec: SOTA Zero Shot Image Captioning Using CLIP and GPT2, EMNLP 2022 (findings)☆209Jan 28, 2024Updated 2 years ago
- [ECCV 2022 Oral] Source code for "A Perturbation-Constrained Adversarial Attack for Evaluating the Robustness of Optical Flow"☆15Dec 13, 2022Updated 3 years ago
- Data repository for the VALSE benchmark.☆40Feb 15, 2024Updated 2 years ago
- Pytorch code for the paper "The color out of space: learning self-supervised representations for Earth Observation imagery"☆18Oct 26, 2021Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- An automatic MLLM hallucination detection framework☆19Sep 26, 2023Updated 2 years ago
- Solving UCF-101 with fastai2☆28Apr 12, 2023Updated 3 years ago
- Some papers about *diverse* image (a few videos) captioning☆25Apr 4, 2023Updated 3 years ago
- PyTorch code for "Fine-grained Image Captioning with CLIP Reward" (Findings of NAACL 2022)☆246Jun 10, 2025Updated last year
- Extend BoxDiff to SDXL (SDXL-based layout-to-image generation)☆28May 23, 2024Updated 2 years ago
- NICE challenge 2023 Track2 2nd result(total 4th) (CVPR 2023) sponsered by LG AI/Shutterstock/SNU☆11Jun 22, 2023Updated 3 years ago
- Example Cloudflare Workers project showing how to return HTML responses with enriched region data☆14May 3, 2021Updated 5 years ago
- Source code for the paper: "Deep Segmentation of the Mandibular Canal: a New 3D Annotated Dataset of CBCT Volumes.", IEEE Access☆24Dec 5, 2022Updated 3 years ago
- ICLR 2023 DeCap: Decoding CLIP Latents for Zero-shot Captioning☆144Mar 16, 2023Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- CVPR-NTIRE 2025 Challenge on UGC Video Enhancement☆23May 30, 2025Updated last year
- Select filter features with mutual-information-based methods☆11Dec 4, 2023Updated 2 years ago
- Deferring loading of JS files until after React loads☆10Dec 4, 2022Updated 3 years ago
- ☆28Sep 22, 2022Updated 3 years ago
- ☆55Aug 3, 2023Updated 3 years ago
- ☆29Feb 23, 2026Updated 6 months ago
- ☆47Oct 5, 2025Updated 10 months ago
- ☆16May 24, 2024Updated 2 years ago
- ☆12Jun 1, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Non-disruptive collagen characterization in clinical histopathology using cross-modality image synthesis☆11Apr 25, 2025Updated last year
- [ICCV2023] Tem-adapter: Adapting Image-Text Pretraining for Video Question Answer☆37Oct 18, 2023Updated 2 years ago
- Learning globally stable dynamical systems policies through imitation. A modification of the original work, focussing on waypoint-based i…☆14Oct 12, 2024Updated last year
- An open-source implementaion for fine-tuning DINOv2 by Meta.☆18Jul 21, 2025Updated last year
- Typper's POC: Plan and discover perfect trips with Phidata and OpenAI using smart prompts.☆16May 17, 2024Updated 2 years ago
- ☆22May 4, 2023Updated 3 years ago
- [ICCV 23] Frequency Guidance Matters in Few-shot Learning☆17Aug 3, 2024Updated 2 years ago