XGEN-MM(BLIP3) Autocaptioning Tools
☆17Jun 20, 2024Updated 2 years ago
Alternatives and similar repositories for Consume-Blip3
Users that are interested in Consume-Blip3 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆13Dec 17, 2024Updated last year
- ☆19Aug 19, 2024Updated last year
- ☆11Feb 14, 2024Updated 2 years ago
- Useful utilities for huggingface☆25Dec 26, 2025Updated 7 months ago
- Official Implementation for "Platypose: Calibrated Zero-Shot Multi-Hypothesis 3D Human Motion Estimation"☆15May 6, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- GDPnet: "Geometry-guided Dense Perspective Network for Speech-Driven Facial Animation." (TVCG 2021)☆11Nov 21, 2021Updated 4 years ago
- A software to automatically tag images. It's primary use is for training Stable Diffusion checkpoints and loras.☆24Dec 4, 2025Updated 8 months ago
- A Focal Transformer for Boundary-aware Prostate Segmentation using CT Images☆11Nov 10, 2024Updated last year
- ☆18Dec 29, 2023Updated 2 years ago
- ☆28Feb 10, 2026Updated 6 months ago
- ☆18Apr 9, 2024Updated 2 years ago
- Multimodal Representation Alignment for Image Generation: Text-Image Interleaved Control Is Easier Than You Think!☆122Mar 4, 2025Updated last year
- A SDXL trainer modified from kohya trainer.☆24Dec 3, 2025Updated 8 months ago
- ☆14Jan 22, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Embroid: Unsupervised Prediction Smoothing Can Improve Few-Shot Classification☆11Aug 12, 2023Updated 3 years ago
- INF-LLaVA: Dual-perspective Perception for High-Resolution Multimodal Large Language Model☆42Aug 4, 2024Updated 2 years ago
- ☆16Feb 18, 2023Updated 3 years ago
- [ICML‘25] Official code for paper "Occult: Optimizing Collaborative Communication across Experts for Accelerated Parallel MoE Training an…☆13Apr 17, 2025Updated last year
- Evaluation for stable diffusion model training☆27Aug 24, 2024Updated last year
- ☆11Feb 26, 2024Updated 2 years ago
- [NeurIPS 2024] Image Understanding Makes for A Good Tokenizer for Image Generation☆21Dec 17, 2024Updated last year
- MambaClinix: Hierarchical Gated Convolution and Mamba-Structured UNet for Enhanced 3D Medical Image Segmentation☆24Sep 20, 2024Updated last year
- ☆111Jun 26, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Expo project template for ReScript☆12Oct 5, 2020Updated 5 years ago
- ☆12Jul 6, 2026Updated last month
- Official Implementation of ACL2023: Don't Parse, Choose Spans! Continuous and Discontinuous Constituency Parsing via Autoregressive Span …☆14Aug 25, 2023Updated 2 years ago
- 自分用のカスタムノード☆15Jun 6, 2026Updated 2 months ago
- A Chrome extension that lets you see through the LinkedIn jargon. Inspired by John Carpenter's They Live.☆10Feb 22, 2018Updated 8 years ago
- Extension for stable diffusion webui to add advance prompt tuning☆10Nov 13, 2022Updated 3 years ago
- 用大模型批量处理数据,现支持各种大模型做OCR,支持通义千问, 月之暗面, 百度飞桨OCR, OpenAI 和LLAVA。Use LLM to generate or clean data for academic use. Support OCR with qwen, m…☆17Sep 15, 2024Updated last year
- ☆10Aug 20, 2025Updated 11 months ago
- Small notebook to preprocess and evaluate images.☆14Nov 11, 2022Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- a.k.a autoMBW-V2☆10Sep 6, 2024Updated last year
- ☆11Dec 21, 2020Updated 5 years ago
- ☆10Nov 28, 2023Updated 2 years ago
- [ICLR 2025] Official PyTorch Implementation for CPE: Concept Pinpoint Eraser for Text-to-image Diffusion Models via Residual Attention Ga…☆13Apr 7, 2025Updated last year
- ☆16Jul 7, 2023Updated 3 years ago
- Official code for the paper "HEXA-MoE: Efficient and Heterogeneous-Aware MoE Acceleration with Zero Computation Redundancy"☆15Mar 6, 2025Updated last year
- Send images to Eagle with PNGinfo from directory. Extension for Stable Diffusion UI by AUTOMATIC1111☆12Dec 13, 2022Updated 3 years ago