[CVPR 2026] π₯π₯ Official Repo of USO: Unified Style and Subject-Driven Generation via Disentangled and Reward Learning
β1,234Sep 12, 2025Updated 11 months ago
Alternatives and similar repositories for USO
Users that are interested in USO are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR 2026] π₯π₯ Official Repo of UMO: Scaling Multi-Identity Consistency for Image Customization via Matching Rewardβ190Sep 15, 2025Updated 11 months ago
- HunyuanImage-2.1: An Efficient Diffusion Model for High-Resolution (2K) Text-to-Image Generationββ675Oct 14, 2025Updated 10 months ago
- [ICCV 2025] π₯π₯ UNO: A Universal Customization Method for Both Single and Multi-Subject Conditioningβ1,361Sep 12, 2025Updated 11 months ago
- [NeurIPS 2025] Official implementation of "XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulatioβ¦β627Oct 22, 2025Updated 9 months ago
- [CVPR2026 π] Stand-In is a lightweight, plug-and-play framework for identity-preserving video generation.β782Updated this week
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ICLR 2026] ChronoEdit: Towards Temporal Reasoning for Image Editing and World Simulationβ703Nov 20, 2025Updated 8 months ago
- FantasyPortrait: Enhancing Multi-Character Portrait Animation with Expression-Augmented Diffusion Transformersβ509Aug 20, 2025Updated 11 months ago
- [ICCV 2025] Enhancing spatial understanding in text-to-Image diffusion modelsβ94Sep 11, 2025Updated 11 months ago
- An inference and training framework for multiple image input in Flux Kontext devβ444Sep 1, 2025Updated 11 months ago
- β349Sep 15, 2025Updated 10 months ago
- The official code implementation of the paper "OmniConsistency: Learning Style-Agnostic Consistency from Paired Stylization Data."β423Jun 8, 2025Updated last year
- Pusa: Thousands Timesteps Video Diffusion Modelβ685Feb 13, 2026Updated 6 months ago
- Phantom: Subject-Consistent Video Generation via Cross-Modal Alignmentβ1,515Sep 11, 2025Updated 11 months ago
- Qwen-Image-Lightning: Speed up Qwen-Image model with distillationβ1,350Jan 1, 2026Updated 7 months ago
- Proton VPN Special Offer - Get 70% off β’ AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioningβ1,282Jan 25, 2026Updated 6 months ago
- [ICML 2026] Official implementation for "DyPE: Dynamic Position Extrapolation for Ultra High Resolution Diffusion".β357May 18, 2026Updated 2 months ago
- Directly Aligning the Full Diffusion Trajectory with Fine-Grained Human Preferenceβ1,278May 11, 2026Updated 3 months ago
- [ICLR 2026] Taming large-scale few-step training with self-adversarial flows! ππ»β538Feb 24, 2026Updated 5 months ago
- β97Nov 6, 2025Updated 9 months ago
- HunyuanImage-3.0: A Powerful Native Multimodal Model for Image Generationβ3,232Jun 23, 2026Updated last month
- [SIGGRAPH Asia 2025] DreamO: A Unified Framework for Image Customizationβ1,651Aug 14, 2025Updated last year
- Qwen-Image is a powerful image generation foundation model capable of complex text rendering and precise image editing.β8,240Feb 10, 2026Updated 6 months ago
- [CVPR 2026 Highlight] High-Quality Text-to-Video Generation with Alpha Channelβ400Apr 9, 2026Updated 4 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer β’ AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Anemoi: A Semi-Centralized Multi-agent Systems Based on Agent-to-Agent Communication MCP server from Coral Protocolβ370Aug 27, 2025Updated 11 months ago
- https://little-misfit.github.io/GRAG-Image-Editing/β119Nov 27, 2025Updated 8 months ago
- Lynx: Towards High-Fidelity Personalized Video Generationβ335Feb 27, 2026Updated 5 months ago
- Tuning-Free 4K Video Generationβ187Aug 7, 2026Updated last week
- An open-source implementation of Whisperβ491Oct 29, 2025Updated 9 months ago
- [ICCV 2025] Official implementations for paper: VACE: All-in-One Video Creation and Editingβ3,916Oct 17, 2025Updated 9 months ago
- Official inference code and LongText-Bench benchmark for our paper X-Omni (https://arxiv.org/pdf/2507.22058).β428Aug 26, 2025Updated 11 months ago
- SteadyDancer: Harmonized and Coherent Human Image Animation with First-Frame Preservationβ643Dec 23, 2025Updated 7 months ago
- Official Implementation of DRA-Ctrl (Dimension-Reduction Attack! Video Generative Models are Experts on Controllable Image Synthesis)β119Aug 15, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generationβ1,229Oct 15, 2025Updated 10 months ago
- A SOTA open-source image editing model, which aims to provide comparable performance against the closed-source models like GPT-4o and Gemβ¦β2,251Apr 29, 2026Updated 3 months ago
- OmniGen2: Exploration to Advanced Multimodal Generation. https://arxiv.org/abs/2506.18871β4,113Mar 20, 2026Updated 4 months ago
- [NeurIPS 2025] Image editing is worth a single LoRA! 0.1% training data for fantastic image editing! Surpasses GPT-4o in ID persistence~ β¦β2,102Dec 19, 2025Updated 7 months ago
- [CVPR'26] ObjectClear: Precise Object and Effect Removal with Adaptive Target-Aware Attentionβ613Aug 3, 2026Updated last week
- β1,046May 14, 2025Updated last year
- UniWorld: High-Resolution Semantic Encoders for Unified Visual Understanding and Generationβ888Dec 23, 2025Updated 7 months ago