Python scripts to use for captioning images with VLMs
☆48Apr 23, 2025Updated last year
Alternatives and similar repositories for VLM-Captioning-Tools
Users that are interested in VLM-Captioning-Tools are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Using CogVLM and CogAgent for image captioning☆15Dec 29, 2023Updated 2 years ago
- OPSTL: Self-supervised Skeleton-based Action Recognition in Occluded Environments☆14Oct 25, 2023Updated 2 years ago
- ☆13Feb 2, 2024Updated 2 years ago
- Custom LORA training on DynamiCrafter☆18Jul 26, 2024Updated 2 years ago
- Extension/Script for Stable Diffusion UI by AUTOMATIC1111 https://github.com/AUTOMATIC1111/stable-diffusion-webui☆17Dec 19, 2022Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Gradio UI for training video models using finetrainers☆35Apr 18, 2025Updated last year
- Data and codes for energy circuit method-based optimal energy flow☆14Sep 17, 2022Updated 3 years ago
- The Official PyTorch implementation of "Part Aware Contrastive Learning for Self-Supervised Action Recognition" in IJCAI 2023☆13Nov 9, 2023Updated 2 years ago
- Condensed Movies Challenge 2021☆24Sep 21, 2022Updated 3 years ago
- ☆10Jun 30, 2026Updated 2 months ago
- XGEN-MM(BLIP3) Autocaptioning Tools☆17Jun 20, 2024Updated 2 years ago
- ComfyUI custom nodes for Diffusion Attentive Attribution Maps (DAAM)☆54Oct 13, 2025Updated 10 months ago
- A funny extension that integrates image-browsing , downloader , deduplicate , cluster , can quickly collect, classify and process your im…☆24Aug 8, 2023Updated 3 years ago
- This repository is the official PyTorch implementation of MA-AGIQA:Large Multi-modality Model Assisted AI-Generated Image Quality Assessm…☆27Sep 28, 2024Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- A multimodal inference pipeline that integrates InstructBLIP with textgen-webui for Vicuna and related models.☆33Jul 14, 2023Updated 3 years ago
- Various training scripts used to train bigasp☆112Aug 13, 2025Updated last year
- This repository focuses on the use of keystroke dynamics as a behavioral biometric to build machine learning models for user recognition.…☆13Mar 29, 2023Updated 3 years ago
- Official impl. of "MagicMirror: A Large-Scale Dataset and Benchmark for Fine-Grained Artifacts Assessment in Text-to-Image Generation"☆24Sep 15, 2025Updated 11 months ago
- The program used to occupy GPUs.☆10Mar 24, 2023Updated 3 years ago
- Snoopy v2.0.1 - Automated digital terrestrial tracking framework☆11Feb 10, 2017Updated 9 years ago
- Virtual sounds for your keystrokes. Compatible with every OS.☆11Apr 3, 2023Updated 3 years ago
- A sd-webui extension for utilizing DanTagGen to "upsample prompts".☆12Jun 13, 2024Updated 2 years ago
- PyTorch based image conversions between equirectangular, cubemap, and perspective. Based on py360convert☆36Aug 21, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- The Florence Tool CLI provides a command-line interface for processing images using the Florence-2 model. This tool allows users to apply…☆16Jan 21, 2025Updated last year
- Fine-tuning code for CLIP models☆274Jun 9, 2026Updated 2 months ago
- attention으로 시계열 예측은 할 수 없을까☆10Apr 30, 2021Updated 5 years ago
- ☆55Jun 24, 2025Updated last year
- ☆15Nov 15, 2023Updated 2 years ago
- Awesome GAN-based Image Restoration☆12Mar 11, 2024Updated 2 years ago
- Official code for the CVPR 2024 Paper "Can Biases in ImageNet Models Explain Generalization?".☆13Jun 24, 2024Updated 2 years ago
- [ICCV 2023] Latent Action Composition for Skeleton-based Action Segmentation☆22Oct 25, 2023Updated 2 years ago
- ☆70Oct 6, 2023Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- 大学Latex答辩模版,当前包含川大、哈工大、中科大。☆11Jul 22, 2024Updated 2 years ago
- DINO-based perceptual losses and FDD feature extraction☆40Jan 7, 2026Updated 7 months ago
- A segmentation project based on aniseg, trained on yolov8-seg☆13Jul 15, 2023Updated 3 years ago
- Official code release for the paper Trapped in texture bias? A large scale comparison of deep instance segmentation, accepted at ECCV 202…☆15Jan 16, 2024Updated 2 years ago
- PEA-Diffusion: Parameter-Efficient Adapter with Knowledge Distillation in non-English Text-to-Image Generation☆37Oct 28, 2024Updated last year
- Official implementation of "Single Image Iterative Subject-driven Generation and Editing".☆99May 30, 2025Updated last year
- ☆14Aug 4, 2023Updated 3 years ago