Multimodal Representation Alignment for Image Generation: Text-Image Interleaved Control Is Easier Than You Think!
☆123Mar 4, 2025Updated last year
Alternatives and similar repositories for DreamEngine
Users that are interested in DreamEngine are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [NAACL 2025] Source code for MMEvalPro, a more trustworthy and efficient benchmark for evaluating LMMs☆25Sep 26, 2024Updated last year
- ☆35Mar 18, 2025Updated last year
- ☆13Jul 10, 2024Updated 2 years ago
- ICML2025, I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models☆191Sep 7, 2025Updated 10 months ago
- ☆292Jul 29, 2025Updated 11 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [CVPR2025] Official Implementation of ILLUME+☆126Aug 20, 2025Updated 11 months ago
- ☆109Nov 27, 2024Updated last year
- [AAAI 2026] This is the official implementation of "T-LoRA: Single Image Diffusion Model Customization Without Overfitting"☆149Apr 24, 2026Updated 2 months ago
- [AAAI 2026] Personalize Anything for Free with Diffusion Transformer☆361Mar 26, 2026Updated 3 months ago
- ☆196Dec 17, 2024Updated last year
- ☆571Nov 26, 2024Updated last year
- RepText: Rendering Visual Text via Replicating 🔥☆139Jun 7, 2025Updated last year
- Official code for ICCV 2025 paper, X2I: Seamless Integration of Multimodal Understanding into Diffusion Transformer via Attention Distill…☆89Jun 26, 2025Updated last year
- Official repository of "GoT: Unleashing Reasoning Capability of Multimodal Large Language Model for Visual Generation and Editing"☆317Sep 28, 2025Updated 9 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICCV 2025] Official repo for "GigaTok: Scaling Visual Tokenizers to 3 Billion Parameters for Autoregressive Image Generation"☆204Jan 7, 2026Updated 6 months ago
- [CVPR2025] PyTorch-based reimplementation of CrossFlow, as proposed in 'Flowing from Words to Pixels: A Noise-Free Framework for Cross-Mo…☆343Jun 8, 2025Updated last year
- ☆10Jul 25, 2023Updated 2 years ago
- A SOTA open-source image editing model, which aims to provide comparable performance against the closed-source models like GPT-4o and Gem…☆2,236Apr 29, 2026Updated 2 months ago
- Official code of "Edit Transfer: Learning Image Editing via Vision In-Context Relations"☆89Jun 6, 2025Updated last year
- [ICCV 2025] Official implementation for KV-Edit: Training-Free Image Editing for Precise Background Preservation☆386May 21, 2025Updated last year
- [ICCV 2025] Code Implementation of "ArtEditor: Learning Customized Instructional Image Editor from Few-Shot Examples"☆430Apr 23, 2025Updated last year
- [ICCV 2025] FreeFlux: Understanding and Exploiting Layer-Specific Roles in RoPE-Based MMDiT for Versatile Image Editing☆77Mar 7, 2026Updated 4 months ago
- [NeurIPS 2025] Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations☆202Sep 18, 2025Updated 10 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Lumina-Image 2.0: A Unified and Efficient Image Generative Framework☆1,006May 22, 2026Updated last month