CLIPxGPT Captioner is Image Captioning Model based on OpenAI's CLIP and GPT-2.
☆118Feb 17, 2025Updated last year
Alternatives and similar repositories for clip-gpt-captioning
Users that are interested in clip-gpt-captioning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV 2023] Going Beyond Nouns With Vision & Language Models Using Synthetic Data☆13Sep 30, 2023Updated 2 years ago
- Differentiable Patch Selection☆15Feb 20, 2023Updated 3 years ago
- An up-to-date & curated list of awesome layout to image papers, methods & resources.☆13Jun 28, 2024Updated 2 years ago
- Simple image captioning model☆1,421Jun 9, 2024Updated 2 years ago
- This repository contains the code and datasets for our ICCV-W paper 'Enhancing CLIP with GPT-4: Harnessing Visual Descriptions as Prompts…☆30Feb 21, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆11May 5, 2024Updated 2 years ago
- finetune script for SDXL adapted from waifu-diffusion trainer☆11Aug 21, 2023Updated 2 years ago
- GRIT: Faster and Better Image-captioning Transformer (ECCV 2022)☆199May 9, 2023Updated 3 years ago
- CapDec: SOTA Zero Shot Image Captioning Using CLIP and GPT2, EMNLP 2022 (findings)☆209Jan 28, 2024Updated 2 years ago
- Pytorch implementation of "ST360IQ: NO-REFERENCE OMNIDIRECTIONAL IMAGE QUALITY ASSESSMENT WITH SPHERICAL VISION TRANSFORMERS"☆15May 12, 2023Updated 3 years ago
- Data repository for the VALSE benchmark.☆40Feb 15, 2024Updated 2 years ago
- ☆13Jul 20, 2024Updated 2 years ago
- Effective training of convolutional neural networks for age estimation based on knowledge distillation☆19Jun 7, 2021Updated 5 years ago
- Code repository for "Improving Detection of Small Oriented Objects in Aerial Images", WACV 2023 MaCVi - Best Paper Award☆14May 6, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆31May 26, 2025Updated last year
- An automatic MLLM hallucination detection framework☆19Sep 26, 2023Updated 2 years ago
- Using a CNN-LSTM hybrid network to generate captions for images☆18Nov 19, 2019Updated 6 years ago
- Jin, Xiao, et al. "FCMNet: Frequency-aware cross-modality attention networks for RGB-D salient object detection." Neurocomputing 491 (202…☆11Apr 11, 2024Updated 2 years ago
- Implementation of 'End-to-End Transformer Based Model for Image Captioning' [AAAI 2022]☆70Jun 1, 2024Updated 2 years ago
- This repository is for the paper "Is BERT Blind? Exploring the Effect of Vision-and-Language Pretraining on Visual Language Understanding…☆21Nov 2, 2023Updated 2 years ago
- Improving neural network representations using human similarity judgments☆13Nov 22, 2024Updated last year
- PyTorch code for "Fine-grained Image Captioning with CLIP Reward" (Findings of NAACL 2022)☆246Jun 10, 2025Updated last year
- NICE challenge 2023 Track2 2nd result(total 4th) (CVPR 2023) sponsered by LG AI/Shutterstock/SNU☆11Jun 22, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ICLR 2023 DeCap: Decoding CLIP Latents for Zero-shot Captioning☆144Mar 16, 2023Updated 3 years ago
- Relative Total Variation(a method for structure extraction from texture)☆20Jul 4, 2024Updated 2 years ago
- This repo contains the code to reproduce the paper: "Enriched Music Representations with Multiple Cross-modal Contrastive Learning"☆15Jun 22, 2023Updated 3 years ago
- [IJCV2025] https://arxiv.org/abs/2304.04521☆16Jan 22, 2025Updated last year
- A large-scale benchmark for the evaluation of embeddings across a number of fine-grained and instance-level visual domains.☆17Jun 14, 2024Updated 2 years ago
- Modality-Agnostic Self-Supervised Learning with Meta-Learned Masked Auto-Encoder (NeurIPS 2023)☆10Jun 5, 2024Updated 2 years ago
- Code for VCRNet: Visual Compensation Restoration Network for No-Reference Image Quality Assessment☆25Apr 12, 2023Updated 3 years ago
- 🔥 Pytorch implementation for a feedback saliency detection model (SalFBNet)☆24Feb 5, 2025Updated last year
- Official code for Deep Bayesian Video Frame Interpolation (ECCV2022)☆18May 29, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- The torchosr module is a set of tools for Open Set Recognition in Python, compatible with PyTorch library.☆13Mar 11, 2025Updated last year
- Code and data for the COLING 2020 paper "Try to Substitute: An Unsupervised Chinese Word Sense Disambiguation Method Based on HowNet"☆14Dec 2, 2020Updated 5 years ago
- ☆22Sep 4, 2023Updated 2 years ago
- ☆14Feb 25, 2023Updated 3 years ago
- ☆16May 24, 2024Updated 2 years ago
- ☆12Jun 1, 2024Updated 2 years ago
- Non-disruptive collagen characterization in clinical histopathology using cross-modality image synthesis☆11Apr 25, 2025Updated last year