☆21Sep 28, 2024Updated last year
Alternatives and similar repositories for qwen2vl-captioner-gui
Users that are interested in qwen2vl-captioner-gui are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Various training scripts used to train bigasp☆113Aug 13, 2025Updated 11 months ago
- ACE-Step: A Step Towards Music Generation Foundation Model☆50May 20, 2025Updated last year
- we generate captions to the images which are given by user(user input) using prompt engineering and Generative AI☆10Aug 24, 2024Updated last year
- A WebUI for Side-by-Side Comparison of Media (Images/Videos) Across Multiple Folders☆26Feb 21, 2025Updated last year
- Easy Pony is a helper node that simplifies the process of adding scoring and other attributes to the core when prompting with Pony models…☆11Apr 5, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Using image caption models to extract prompts in ComfyUI☆12May 21, 2025Updated last year
- QLoRA: Efficient Finetuning of Quantized LLMs☆11Jul 22, 2023Updated 3 years ago
- The Florence Tool CLI provides a command-line interface for processing images using the Florence-2 model. This tool allows users to apply…☆16Jan 21, 2025Updated last year
- ComfyUI implementation of Motion-I2V☆41Sep 30, 2024Updated last year
- Extract dominant or complementary color palettes from images. Convert colors to English names suitable for txt2img prompts.☆16Jan 5, 2025Updated last year
- An easy-to-use GUI addon for whisper-standalone-win. Designed for those who prefer a simple interface over typing commands and file paths…☆13Dec 26, 2023Updated 2 years ago
- Extract individual frames from a video as png images (android)☆13Dec 30, 2022Updated 3 years ago
- Open Translator: Speech To Speech and Speech to text Translator with voice cloning and other cool features☆18Mar 26, 2026Updated 4 months ago
- implementation of https://arxiv.org/pdf/2312.09299☆21Jul 3, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Image caption and manage tool for AI training☆11Jan 24, 2025Updated last year
- Windows ComfyUI Installer GUI☆11Mar 30, 2025Updated last year
- 使用mnn-llm对GOT-OCR2.0进行推理☆14Oct 2, 2024Updated last year
- A program to help with Ideogram captioning☆30Updated this week
- ☆13Feb 5, 2026Updated 5 months ago
- A real time offline transcriber with gui, based on OpenAI whisper☆17Dec 25, 2025Updated 7 months ago
- ComfyUI Yolo World EfficientSAM custom node☆15Jul 16, 2024Updated 2 years ago
- ☆14Oct 31, 2023Updated 2 years ago
- ☆15May 26, 2026Updated 2 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A batch captioning tool for joy_caption☆196Aug 25, 2025Updated 11 months ago
- A powerful and user-friendly tool that generates detailed captions for your images☆21Nov 11, 2024Updated last year
- Flux Fill 1.0 GO: flux Inpainting and outpainting starting with 8Gb of VRAM☆78Jan 18, 2025Updated last year
- An AI try-on application for generating photos with AI character wearing the same clothes as the one in the input photo.☆14Sep 7, 2023Updated 2 years ago
- A Python-based chatbot project built on the autogen and tinygrad foundation, utilizing advanced agents for dynamic conversations and func…☆27Oct 9, 2024Updated last year
- [AAAI 2025] Official Implementation of 3D$^2$-Actor: Learning Pose-Conditioned 3D-Aware Denoiser for Realistic Gaussian Avatar Modeling☆15Mar 30, 2025Updated last year
- ☆17Nov 6, 2025Updated 8 months ago
- Koel Labs innovates open-source speech research, inclusive speech technologies, and real-time pronunciation feedback for language learner…☆14Jul 13, 2026Updated 2 weeks ago
- Official implementation of "iRBSM: A Deep Implicit 3D Breast Shape Model" (BVM'25).☆15Dec 2, 2025Updated 7 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A ComfyUI node for describing an image☆20May 22, 2024Updated 2 years ago
- A pipeline to generate user-preferred photo-realistic avatars using stable-diffusion and bayesian-optimization.☆18Jun 12, 2026Updated last month
- Automatic1111 port of my comfyUI geely remb tool☆17Oct 24, 2024Updated last year
- A fork of the PEFT library, supporting Robust Adaptation (RoSA)☆15Aug 16, 2024Updated last year
- Official code for CVPR2022 paper: Depth-Aware Generative Adversarial Network for Talking Head Video Generation☆23Sep 14, 2022Updated 3 years ago
- body reconstruction☆16Jun 23, 2021Updated 5 years ago
- SYDE 671 final project source code, copied from Google Research to avoid cloning the entire Google Research repo.☆14Nov 17, 2019Updated 6 years ago