[CVPR 2026] π₯π₯ Official Repo of USO: Unified Style and Subject-Driven Generation via Disentangled and Reward Learning
β1,236Sep 12, 2025Updated last year
Alternatives and similar repositories for USO
Users that are interested in USO are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR 2026] π₯π₯ Official Repo of UMO: Scaling Multi-Identity Consistency for Image Customization via Matching Rewardβ190Sep 15, 2025Updated last year
- HunyuanImage-2.1: An Efficient Diffusion Model for High-Resolution (2K) Text-to-Image Generationββ674Oct 14, 2025Updated 11 months ago
- [ICCV 2025] π₯π₯ UNO: A Universal Customization Method for Both Single and Multi-Subject Conditioningβ1,362Sep 12, 2025Updated last year
- [NeurIPS 2025] Official implementation of "XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulatioβ¦β627Oct 22, 2025Updated 11 months ago
- [CVPR2026 π] Stand-In is a lightweight, plug-and-play framework for identity-preserving video generation.β791Aug 10, 2026Updated last month
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICLR 2026] ChronoEdit: Towards Temporal Reasoning for Image Editing and World Simulationβ707Nov 20, 2025Updated 10 months ago
- FantasyPortrait: Enhancing Multi-Character Portrait Animation with Expression-Augmented Diffusion Transformersβ511Aug 20, 2025Updated last year
- [ICCV 2025] Enhancing spatial understanding in text-to-Image diffusion modelsβ95Sep 11, 2025Updated last year
- An inference and training framework for multiple image input in Flux Kontext devβ443Sep 1, 2025Updated last year
- β349Sep 15, 2025Updated last year
- The official code implementation of the paper "OmniConsistency: Learning Style-Agnostic Consistency from Paired Stylization Data."β423Jun 8, 2025Updated last year
- Pusa: Thousands Timesteps Video Diffusion Modelβ686Feb 13, 2026Updated 7 months ago
- Phantom: Subject-Consistent Video Generation via Cross-Modal Alignmentβ1,517Sep 11, 2025Updated last year
- Qwen-Image-Lightning: Speed up Qwen-Image model with distillationβ1,361Aug 31, 2026Updated 3 weeks ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioningβ1,286Jan 25, 2026Updated 7 months ago
- [ICML 2026] Official implementation for "DyPE: Dynamic Position Extrapolation for Ultra High Resolution Diffusion".β360May 18, 2026Updated 4 months ago
- Directly Aligning the Full Diffusion Trajectory with Fine-Grained Human Preferenceβ1,279May 11, 2026Updated 4 months ago
- [ICLR 2026] Taming large-scale few-step training with self-adversarial flows! ππ»β541Feb 24, 2026Updated 7 months ago
- β97Nov 6, 2025Updated 10 months ago
- HunyuanImage-3.0: A Powerful Native Multimodal Model for Image Generationβ3,277Jun 23, 2026Updated 3 months ago
- [SIGGRAPH Asia 2025] DreamO: A Unified Framework for Image Customizationβ1,648Aug 14, 2025Updated last year
- Qwen-Image is a powerful image generation foundation model capable of complex text rendering and precise image editing.β8,367Feb 10, 2026Updated 7 months ago
- [CVPR 2026 Highlight] High-Quality Text-to-Video Generation with Alpha Channelβ419Apr 9, 2026Updated 5 months ago
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Anemoi: A Semi-Centralized Multi-agent Systems Based on Agent-to-Agent Communication MCP server from Coral Protocolβ370Aug 27, 2025Updated last year
- https://little-misfit.github.io/GRAG-Image-Editing/β119Nov 27, 2025Updated 9 months ago
- Lynx: Towards High-Fidelity Personalized Video Generationβ335Feb 27, 2026Updated 6 months ago
- Tuning-Free 4K Video Generationβ213Updated this week
- An open-source implementation of Whisperβ493Oct 29, 2025Updated 10 months ago
- [ICCV 2025] Official implementations for paper: VACE: All-in-One Video Creation and Editingβ3,953Oct 17, 2025Updated 11 months ago
- Official inference code and LongText-Bench benchmark for our paper X-Omni (https://arxiv.org/pdf/2507.22058).β428Aug 26, 2025Updated last year
- SteadyDancer: Harmonized and Coherent Human Image Animation with First-Frame Preservationβ648Dec 23, 2025Updated 9 months ago
- Official Implementation of DRA-Ctrl (Dimension-Reduction Attack! Video Generative Models are Experts on Controllable Image Synthesis)β119Aug 15, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generationβ1,230Oct 15, 2025Updated 11 months ago
- A SOTA open-source image editing model, which aims to provide comparable performance against the closed-source models like GPT-4o and Gemβ¦β2,264Apr 29, 2026Updated 4 months ago
- OmniGen2: Exploration to Advanced Multimodal Generation. https://arxiv.org/abs/2506.18871β4,113Mar 20, 2026Updated 6 months ago
- [NeurIPS 2025] Image editing is worth a single LoRA! 0.1% training data for fantastic image editing! Surpasses GPT-4o in ID persistence~ β¦β2,105Dec 19, 2025Updated 9 months ago
- [CVPR'26] ObjectClear: Precise Object and Effect Removal with Adaptive Target-Aware Attentionβ623Aug 3, 2026Updated last month
- β1,043May 14, 2025Updated last year
- UniWorld: High-Resolution Semantic Encoders for Unified Visual Understanding and Generationβ891Dec 23, 2025Updated 9 months ago