Python scripts to use for captioning images with VLMs
☆47Apr 23, 2025Updated last year
Alternatives and similar repositories for VLM-Captioning-Tools
Users that are interested in VLM-Captioning-Tools are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- OPSTL: Self-supervised Skeleton-based Action Recognition in Occluded Environments☆14Oct 25, 2023Updated 2 years ago
- ☆13Feb 2, 2024Updated 2 years ago
- Custom LORA training on DynamiCrafter☆18Jul 26, 2024Updated 2 years ago
- Extension/Script for Stable Diffusion UI by AUTOMATIC1111 https://github.com/AUTOMATIC1111/stable-diffusion-webui☆17Dec 19, 2022Updated 3 years ago
- Gradio UI for training video models using finetrainers☆35Apr 18, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The Official PyTorch implementation of "Part Aware Contrastive Learning for Self-Supervised Action Recognition" in IJCAI 2023☆13Nov 9, 2023Updated 2 years ago
- A tool to visualize convolutional layer activations on an input image.☆12Jun 5, 2016Updated 10 years ago
- XGEN-MM(BLIP3) Autocaptioning Tools☆17Jun 20, 2024Updated 2 years ago
- A funny extension that integrates image-browsing , downloader , deduplicate , cluster , can quickly collect, classify and process your im…☆24Aug 8, 2023Updated 2 years ago
- This repository is the official PyTorch implementation of MA-AGIQA:Large Multi-modality Model Assisted AI-Generated Image Quality Assessm…☆27Sep 28, 2024Updated last year
- A multimodal inference pipeline that integrates InstructBLIP with textgen-webui for Vicuna and related models.☆33Jul 14, 2023Updated 3 years ago
- Official impl. of "MagicMirror: A Large-Scale Dataset and Benchmark for Fine-Grained Artifacts Assessment in Text-to-Image Generation"☆24Sep 15, 2025Updated 10 months ago
- The program used to occupy GPUs.☆10Mar 24, 2023Updated 3 years ago
- Snoopy v2.0.1 - Automated digital terrestrial tracking framework☆11Feb 10, 2017Updated 9 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Virtual sounds for your keystrokes. Compatible with every OS.☆11Apr 3, 2023Updated 3 years ago
- A sd-webui extension for utilizing DanTagGen to "upsample prompts".☆12Jun 13, 2024Updated 2 years ago
- PyTorch based image conversions between equirectangular, cubemap, and perspective. Based on py360convert☆35Aug 21, 2025Updated 11 months ago
- Simple PHP Script to return your true external ip (wan)☆11Mar 7, 2015Updated 11 years ago
- The Florence Tool CLI provides a command-line interface for processing images using the Florence-2 model. This tool allows users to apply…☆16Jan 21, 2025Updated last year
- Running Arduino sketches on Mbed OS☆15Aug 9, 2019Updated 6 years ago
- ✨ waifu-diffusion tagger server / onnx | wd-tagger as api service☆21Feb 20, 2025Updated last year
- Fine-tuning code for CLIP models☆275Jun 9, 2026Updated last month
- ☆15Nov 15, 2023Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Official code for the CVPR 2024 Paper "Can Biases in ImageNet Models Explain Generalization?".☆13Jun 24, 2024Updated 2 years ago
- ☆70Oct 6, 2023Updated 2 years ago
- DINO-based perceptual losses and FDD feature extraction☆36Jan 7, 2026Updated 6 months ago
- Official code release for the paper Trapped in texture bias? A large scale comparison of deep instance segmentation, accepted at ECCV 202…☆16Jan 16, 2024Updated 2 years ago
- Python tool for creating average images from faces☆14Jan 13, 2017Updated 9 years ago
- PEA-Diffusion: Parameter-Efficient Adapter with Knowledge Distillation in non-English Text-to-Image Generation☆37Oct 28, 2024Updated last year
- sample microservice project to demonstrate use of Azure keyvault and Kubernetes ConfigMaps for Configuration, use of Serilog for strucute…☆11Feb 6, 2018Updated 8 years ago
- Pytorch Sketch Classification☆11Apr 14, 2018Updated 8 years ago
- Official implementation of "Single Image Iterative Subject-driven Generation and Editing".☆99May 30, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆21Sep 28, 2024Updated last year
- Useful utilities for huggingface☆25Dec 26, 2025Updated 7 months ago
- A large scale dataset for Video Captioning in Italian☆13May 16, 2023Updated 3 years ago
- [arXiv 2026] This is the official PyTorch implementation of "RTDMD: Reinforcing Few-step Generators via Reward-Tilted Distribution Matchi…☆41Jun 6, 2026Updated last month
- Model code for inferencing T5☆67Mar 10, 2025Updated last year
- Collect papers and codes about VQGAN in various Computer Vision tasks☆10Dec 20, 2022Updated 3 years ago
- Beyond Textual CoT: Interleaved Text-image chains with Deep Confidence Reasoning for Image Editing☆19Jun 24, 2026Updated last month