Code and models for the paper "One Transformer Fits All Distributions in Multi-Modal Diffusion"
☆1,486May 31, 2023Updated 3 years ago
Alternatives and similar repositories for unidiffuser
Users that are interested in unidiffuser are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A PyTorch implementation of the paper "All are Worth Words: A ViT Backbone for Diffusion Models".☆1,110Mar 25, 2023Updated 3 years ago
- Masked Diffusion Transformer is the SOTA for image synthesis. (ICCV 2023)☆596Apr 23, 2024Updated 2 years ago
- Versatile Diffusion: Text, Images and Variations All in One Diffusion Model, arXiv 2022 / ICCV 2023☆1,334Aug 10, 2023Updated 3 years ago
- Official repo for consistency models.☆6,486Mar 22, 2024Updated 2 years ago
- Official PyTorch Implementation of "Scalable Diffusion Models with Transformers"☆8,690May 31, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Consistency Distilled Diff VAE☆2,212Nov 7, 2023Updated 2 years ago
- T2I-Adapter☆3,801Jun 21, 2024Updated 2 years ago
- PixArt-α: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis☆3,304Oct 31, 2024Updated last year
- Official code for "DPM-Solver: A Fast ODE Solver for Diffusion Probabilistic Model Sampling in Around 10 Steps" (Neurips 2022 Oral)☆1,852Feb 6, 2024Updated 2 years ago
- ☆3,456May 14, 2024Updated 2 years ago
- Emu Series: Generative Multimodal Models from BAAI☆1,778Jan 12, 2026Updated 7 months ago
- Official PyTorch Implementation of "SiT: Exploring Flow and Diffusion-based Generative Models with Scalable Interpolant Transformers"☆1,206Dec 22, 2025Updated 8 months ago
- Autoregressive Model Beats Diffusion: 🦙 Llama for Scalable Image Generation☆1,966Aug 15, 2024Updated 2 years ago
- [IJCV] FastComposer: Tuning-Free Multi-Subject Image Generation with Localized Attention☆715Jan 10, 2025Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation (ICCV 2023, Oral)☆541Jan 8, 2024Updated 2 years ago
- [SIGGRAPH Asia 2024] ReVersion: Diffusion-Based Relation Inversion from Images☆503Oct 7, 2025Updated 10 months ago
- Official implementation of "Composer: Creative and Controllable Image Synthesis with Composable Conditions"☆1,557Dec 26, 2023Updated 2 years ago
- Official PyTorch implementation of the paper "In-Context Learning Unlocked for Diffusion Models"☆413Mar 25, 2024Updated 2 years ago
- LAVIS - A One-stop Library for Language-Vision Intelligence☆11,262Jun 2, 2026Updated 2 months ago
- ☆7,419Jul 2, 2024Updated 2 years ago
- Open-Set Grounded Text-to-Image Generation☆2,223Mar 6, 2024Updated 2 years ago
- Implementation of "SVDiff: Compact Parameter Space for Diffusion Fine-Tuning"☆385Jan 24, 2024Updated 2 years ago
- Custom Diffusion: Multi-Concept Customization of Text-to-Image Diffusion (CVPR 2023)☆1,978May 24, 2026Updated 3 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference☆4,615Jun 14, 2024Updated 2 years ago
- [CVPR 2024] PAIR Diffusion: A Comprehensive Multimodal Object-Level Image Editor☆522Apr 2, 2024Updated 2 years ago
- [ICML 2024] Mastering Text-to-Image Diffusion: Recaptioning, Planning, and Generating with Multimodal LLMs (RPG)☆1,843Feb 1, 2025Updated last year
- Official Pytorch Implementation for "MultiDiffusion: Fusing Diffusion Paths for Controlled Image Generation" presenting "MultiDiffusion" …☆1,066Sep 21, 2023Updated 2 years ago
- VideoSys: An easy and efficient system for video generation☆2,022Aug 27, 2025Updated last year
- Speed up Stable Diffusion with this one simple trick!☆1,404Nov 29, 2023Updated 2 years ago
- Official implementation of SEED-LLaMA (ICLR 2024).☆642Sep 21, 2024Updated last year
- VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models☆5,074Jan 9, 2026Updated 7 months ago
- [NeurIPS 2023] Uni-ControlNet: All-in-One Control to Text-to-Image Diffusion Models☆668Jul 17, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- High-Resolution Image Synthesis with Latent Diffusion Models☆14,132Feb 29, 2024Updated 2 years ago
- Official JAX implementation of MAGVIT: Masked Generative Video Transformer☆1,001Jan 17, 2024Updated 2 years ago
- Using Low-rank adaptation to quickly fine-tune diffusion models.☆7,551Mar 22, 2024Updated 2 years ago
- SEED-Voken: A Series of Powerful Visual Tokenizers☆1,021Nov 25, 2025Updated 9 months ago
- [NeurIPS 2023] ImageReward: Learning and Evaluating Human Preferences for Text-to-image Generation☆1,702Oct 29, 2025Updated 10 months ago
- Lumina-T2X is a unified framework for Text to Any Modality Generation☆2,250Feb 16, 2025Updated last year
- Code for Fast Training of Diffusion Models with Masked Transformers☆431May 15, 2024Updated 2 years ago