[ICCV 2025] Region-Aware Text-to-Image Generation via Hard Binding and Soft Refinement 🔥
☆622Dec 12, 2025Updated 7 months ago
Alternatives and similar repositories for RAG-Diffusion
Users that are interested in RAG-Diffusion are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Training-free Regional Prompting for Diffusion Transformers 🔥☆696Nov 28, 2024Updated last year
- [CVPR 2025] InstanceCap: Improving Text-to-Video Generation via Instance-aware Structured Caption 🔍☆45Jul 5, 2025Updated last year
- CoDi:Subject-Consistent and Pose-Diverse Text-to-Image Generation☆36Aug 1, 2025Updated 11 months ago
- Official repository of In-Context LoRA for Diffusion Transformers☆2,078Dec 20, 2024Updated last year
- [ICCV 2025 Highlight] OminiControl: Minimal and Universal Control for Diffusion Transformer☆1,925Jul 2, 2026Updated 2 weeks ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- L2P: Unlocking Latent Potential for Pixel Generation☆39May 22, 2026Updated last month
- Latest Advances on Autoregressive Visual Models.📖☆28Mar 15, 2025Updated last year
- TextCrafter: Accurately Rendering Multiple Texts in Complex Visual Scenes☆97Nov 26, 2025Updated 7 months ago
- [TPAMI 2026] ConsistentID : Portrait Generation with Multimodal Fine-Grained Identity Preserving☆1,027Jan 2, 2026Updated 6 months ago
- Code for "L2P: Unlocking Latent Potential for Pixel Generation"☆179Jul 11, 2026Updated last week
- [🚀ICML 2025] "Taming Rectified Flow for Inversion and Editing" Using FLUX and HunyuanVideo for image and video editing!☆637May 1, 2025Updated last year
- It is an Android-based application that enables managing hotspot properties through a web interface, providing mobile routing functionali…☆156Jul 14, 2026Updated last week
- User Identity Scaffolding for Multiple OIDC Authentications for User☆95Jun 14, 2025Updated last year
- Official implementation of the paper: "FlowEdit: Inversion-Free Text-Based Editing Using Pre-Trained Flow Models"☆1,009May 27, 2026Updated last month
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICML 2024] Mastering Text-to-Image Diffusion: Recaptioning, Planning, and Generating with Multimodal LLMs (RPG)☆1,838Feb 1, 2025Updated last year
- [CVPR 2025 Highlight🔥] Identity-Preserving Text-to-Video Generation by Frequency Decomposition☆848Apr 14, 2026Updated 3 months ago
- This is the official repository of UltraHR-100K.☆45Nov 21, 2025Updated 8 months ago
- ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment☆1,285Jul 17, 2024Updated 2 years ago
- [CVPR 2025] Hallo3: Highly Dynamic and Realistic Portrait Image Animation with Video Diffusion Transformer☆1,394Mar 13, 2025Updated last year
- [ECCV 2024] The official implementation of paper "BrushNet: A Plug-and-Play Image Inpainting Model with Decomposed Dual-Branch Diffusion"☆1,737Dec 17, 2024Updated last year
- ☆246Nov 24, 2024Updated last year
- Code for SCIS-2025 Paper "UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation".☆1,190Apr 15, 2025Updated last year
- kight is a static analysis tool for c/c++ programs.☆213Dec 27, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICML 2023 Oral, NeurIPS 2023] Official implementations for paper: Customizable Image Synthesis with Multiple Subjects☆446Sep 12, 2023Updated 2 years ago
- [ICCV 2025] 🔥🔥 UNO: A Universal Customization Method for Both Single and Multi-Subject Conditioning☆1,359Sep 12, 2025Updated 10 months ago
- [AAAI 2026] Personalize Anything for Free with Diffusion Transformer☆361Mar 26, 2026Updated 3 months ago
- Evaluation of Text-to-Video Generation Models: A Dynamics Perspective[NeurIPS 2024].☆274Dec 3, 2024Updated last year
- Efficient DiT architecture for text2any tasks, ICLR2025☆446May 10, 2025Updated last year
- Advanced Unsupervised Image Enhancement with GAN☆247Nov 11, 2024Updated last year
- Official Implementation of AttentionShift: Iteratively Estimated Part-based Attention Map for Pointly Supervised Instance Segmentation☆155Oct 18, 2024Updated last year
- Unofficial Implementation of ReplaceAnything: https://aigcdesigngroup.github.io/replace-anything/☆399May 27, 2024Updated 2 years ago
- Allegro is a powerful text-to-video model that generates high-quality videos up to 6 seconds at 15 FPS and 720p resolution from simple te…☆1,133Feb 7, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- An open-source library with a powerful Contrastive Language-and-Motion (CLaM) pre-training evaluator☆99Nov 23, 2025Updated 7 months ago
- ☆251Feb 11, 2025Updated last year
- Welcome to the 'Open-Alteryx-Macro' project. This project is aimed at providing an open-source solution for managing and updating Alteryx…☆156May 25, 2024Updated 2 years ago
- A curated list of papers, code and resources pertaining to image composition/compositing or object/subject insertion/addition/compositing…☆540Apr 30, 2026Updated 2 months ago
- [AAAI 2025] Exploiting Multimodal Spatial-temporal Patterns for Video Object Tracking☆118May 18, 2025Updated last year
- [ICLR 2025] Hallo2: Long-Duration and High-Resolution Audio-driven Portrait Image Animation☆3,722Feb 27, 2025Updated last year
- Code for paper "Self-Taught Recognizer: Toward Unsupervised Adaptation for Speech Foundation Models"☆241May 24, 2024Updated 2 years ago