[CVPR 2026] π₯π₯ Official Repo of USO: Unified Style and Subject-Driven Generation via Disentangled and Reward Learning
β1,236Sep 12, 2025Updated 11 months ago
Alternatives and similar repositories for USO
Users that are interested in USO are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR 2026] π₯π₯ Official Repo of UMO: Scaling Multi-Identity Consistency for Image Customization via Matching Rewardβ190Sep 15, 2025Updated 11 months ago
- HunyuanImage-2.1: An Efficient Diffusion Model for High-Resolution (2K) Text-to-Image Generationββ675Oct 14, 2025Updated 10 months ago
- [ICCV 2025] π₯π₯ UNO: A Universal Customization Method for Both Single and Multi-Subject Conditioningβ1,362Sep 12, 2025Updated 11 months ago
- [NeurIPS 2025] Official implementation of "XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulatioβ¦β627Oct 22, 2025Updated 10 months ago
- [CVPR2026 π] Stand-In is a lightweight, plug-and-play framework for identity-preserving video generation.β786Aug 10, 2026Updated 3 weeks ago
- End-to-end encrypted email - Proton Mail β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- [ICLR 2026] ChronoEdit: Towards Temporal Reasoning for Image Editing and World Simulationβ704Nov 20, 2025Updated 9 months ago
- FantasyPortrait: Enhancing Multi-Character Portrait Animation with Expression-Augmented Diffusion Transformersβ507Aug 20, 2025Updated last year
- [ICCV 2025] Enhancing spatial understanding in text-to-Image diffusion modelsβ94Sep 11, 2025Updated 11 months ago
- An inference and training framework for multiple image input in Flux Kontext devβ444Sep 1, 2025Updated last year
- β349Sep 15, 2025Updated 11 months ago
- The official code implementation of the paper "OmniConsistency: Learning Style-Agnostic Consistency from Paired Stylization Data."β423Jun 8, 2025Updated last year
- Pusa: Thousands Timesteps Video Diffusion Modelβ686Feb 13, 2026Updated 6 months ago
- Phantom: Subject-Consistent Video Generation via Cross-Modal Alignmentβ1,515Sep 11, 2025Updated 11 months ago
- Qwen-Image-Lightning: Speed up Qwen-Image model with distillationβ1,358Updated this week
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioningβ1,285Jan 25, 2026Updated 7 months ago
- [ICML 2026] Official implementation for "DyPE: Dynamic Position Extrapolation for Ultra High Resolution Diffusion".β360May 18, 2026Updated 3 months ago
- Directly Aligning the Full Diffusion Trajectory with Fine-Grained Human Preferenceβ1,278May 11, 2026Updated 3 months ago
- [ICLR 2026] Taming large-scale few-step training with self-adversarial flows! ππ»β538Feb 24, 2026Updated 6 months ago
- β97Nov 6, 2025Updated 9 months ago
- HunyuanImage-3.0: A Powerful Native Multimodal Model for Image Generationβ3,256Jun 23, 2026Updated 2 months ago
- [SIGGRAPH Asia 2025] DreamO: A Unified Framework for Image Customizationβ1,649Aug 14, 2025Updated last year
- Qwen-Image is a powerful image generation foundation model capable of complex text rendering and precise image editing.β8,286Feb 10, 2026Updated 6 months ago
- [CVPR 2026 Highlight] High-Quality Text-to-Video Generation with Alpha Channelβ408Apr 9, 2026Updated 4 months ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Anemoi: A Semi-Centralized Multi-agent Systems Based on Agent-to-Agent Communication MCP server from Coral Protocolβ370Aug 27, 2025Updated last year
- https://little-misfit.github.io/GRAG-Image-Editing/β119Nov 27, 2025Updated 9 months ago
- Lynx: Towards High-Fidelity Personalized Video Generationβ335Feb 27, 2026Updated 6 months ago
- Tuning-Free 4K Video Generationβ188Updated this week
- An open-source implementation of Whisperβ493Oct 29, 2025Updated 10 months ago
- [ICCV 2025] Official implementations for paper: VACE: All-in-One Video Creation and Editingβ3,936Oct 17, 2025Updated 10 months ago
- Official inference code and LongText-Bench benchmark for our paper X-Omni (https://arxiv.org/pdf/2507.22058).β428Aug 26, 2025Updated last year
- SteadyDancer: Harmonized and Coherent Human Image Animation with First-Frame Preservationβ647Dec 23, 2025Updated 8 months ago
- Official Implementation of DRA-Ctrl (Dimension-Reduction Attack! Video Generative Models are Experts on Controllable Image Synthesis)β119Aug 15, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generationβ1,228Oct 15, 2025Updated 10 months ago
- A SOTA open-source image editing model, which aims to provide comparable performance against the closed-source models like GPT-4o and Gemβ¦β2,258Apr 29, 2026Updated 4 months ago
- OmniGen2: Exploration to Advanced Multimodal Generation. https://arxiv.org/abs/2506.18871β4,112Mar 20, 2026Updated 5 months ago
- [NeurIPS 2025] Image editing is worth a single LoRA! 0.1% training data for fantastic image editing! Surpasses GPT-4o in ID persistence~ β¦β2,104Dec 19, 2025Updated 8 months ago
- [CVPR'26] ObjectClear: Precise Object and Effect Removal with Adaptive Target-Aware Attentionβ616Aug 3, 2026Updated last month
- β1,044May 14, 2025Updated last year
- UniWorld: High-Resolution Semantic Encoders for Unified Visual Understanding and Generationβ890Dec 23, 2025Updated 8 months ago