[CVPR 2026] π₯π₯ Official Repo of USO: Unified Style and Subject-Driven Generation via Disentangled and Reward Learning
β1,227Sep 12, 2025Updated 10 months ago
Alternatives and similar repositories for USO
Users that are interested in USO are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR 2026] π₯π₯ Official Repo of UMO: Scaling Multi-Identity Consistency for Image Customization via Matching Rewardβ190Sep 15, 2025Updated 10 months ago
- HunyuanImage-2.1: An Efficient Diffusion Model for High-Resolution (2K) Text-to-Image Generationββ673Oct 14, 2025Updated 9 months ago
- [ICCV 2025] π₯π₯ UNO: A Universal Customization Method for Both Single and Multi-Subject Conditioningβ1,360Sep 12, 2025Updated 10 months ago
- [NeurIPS 2025] Official implementation of "XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulatioβ¦β627Oct 22, 2025Updated 9 months ago
- [CVPR2026 π] Stand-In is a lightweight, plug-and-play framework for identity-preserving video generation.β777Feb 21, 2026Updated 5 months ago
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [ICLR 2026] ChronoEdit: Towards Temporal Reasoning for Image Editing and World Simulationβ698Nov 20, 2025Updated 8 months ago
- FantasyPortrait: Enhancing Multi-Character Portrait Animation with Expression-Augmented Diffusion Transformersβ511Aug 20, 2025Updated 11 months ago
- [ICCV 2025] Enhancing spatial understanding in text-to-Image diffusion modelsβ94Sep 11, 2025Updated 10 months ago
- An inference and training framework for multiple image input in Flux Kontext devβ441Sep 1, 2025Updated 10 months ago
- The official code implementation of the paper "OmniConsistency: Learning Style-Agnostic Consistency from Paired Stylization Data."β423Jun 8, 2025Updated last year
- β348Sep 15, 2025Updated 10 months ago
- Pusa: Thousands Timesteps Video Diffusion Modelβ686Feb 13, 2026Updated 5 months ago
- Phantom: Subject-Consistent Video Generation via Cross-Modal Alignmentβ1,512Sep 11, 2025Updated 10 months ago
- Qwen-Image-Lightning: Speed up Qwen-Image model with distillationβ1,340Jan 1, 2026Updated 6 months ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioningβ1,274Jan 25, 2026Updated 6 months ago
- [ICML 2026] Official implementation for "DyPE: Dynamic Position Extrapolation for Ultra High Resolution Diffusion".β356May 18, 2026Updated 2 months ago
- Directly Aligning the Full Diffusion Trajectory with Fine-Grained Human Preferenceβ1,278May 11, 2026Updated 2 months ago
- [ICLR 2026] Taming large-scale few-step training with self-adversarial flows! ππ»β536Feb 24, 2026Updated 5 months ago
- β97Nov 6, 2025Updated 8 months ago
- HunyuanImage-3.0: A Powerful Native Multimodal Model for Image Generationβ3,200Jun 23, 2026Updated last month
- [SIGGRAPH Asia 2025] DreamO: A Unified Framework for Image Customizationβ1,652Aug 14, 2025Updated 11 months ago
- Qwen-Image is a powerful image generation foundation model capable of complex text rendering and precise image editing.β8,167Feb 10, 2026Updated 5 months ago
- [CVPR 2026 Highlight] High-Quality Text-to-Video Generation with Alpha Channelβ392Apr 9, 2026Updated 3 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI β’ AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Anemoi: A Semi-Centralized Multi-agent Systems Based on Agent-to-Agent Communication MCP server from Coral Protocolβ370Aug 27, 2025Updated 10 months ago
- https://little-misfit.github.io/GRAG-Image-Editing/β119Nov 27, 2025Updated 7 months ago
- Lynx: Towards High-Fidelity Personalized Video Generationβ336Feb 27, 2026Updated 4 months ago
- Tuning-Free 4K Video Generationβ186Updated this week
- An open-source implementation of Whisperβ492Oct 29, 2025Updated 8 months ago
- [ICCV 2025] Official implementations for paper: VACE: All-in-One Video Creation and Editingβ3,883Oct 17, 2025Updated 9 months ago
- Official inference code and LongText-Bench benchmark for our paper X-Omni (https://arxiv.org/pdf/2507.22058).β426Aug 26, 2025Updated 11 months ago
- SteadyDancer: Harmonized and Coherent Human Image Animation with First-Frame Preservationβ638Dec 23, 2025Updated 7 months ago
- Official Implementation of DRA-Ctrl (Dimension-Reduction Attack! Video Generative Models are Experts on Controllable Image Synthesis)β119Aug 15, 2025Updated 11 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI β’ AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generationβ1,226Oct 15, 2025Updated 9 months ago
- A SOTA open-source image editing model, which aims to provide comparable performance against the closed-source models like GPT-4o and Gemβ¦β2,238Apr 29, 2026Updated 2 months ago
- OmniGen2: Exploration to Advanced Multimodal Generation. https://arxiv.org/abs/2506.18871β4,107Mar 20, 2026Updated 4 months ago
- [NeurIPS 2025] Image editing is worth a single LoRA! 0.1% training data for fantastic image editing! Surpasses GPT-4o in ID persistence~ β¦β2,101Dec 19, 2025Updated 7 months ago
- [CVPR'26] ObjectClear: Precise Object and Effect Removal with Adaptive Target-Aware Attentionβ606Feb 26, 2026Updated 5 months ago
- β1,047May 14, 2025Updated last year
- UniWorld: High-Resolution Semantic Encoders for Unified Visual Understanding and Generationβ884Dec 23, 2025Updated 7 months ago