Official code of "MakeAnything: Harnessing Diffusion Transformers for Multi-Domain Procedural Sequence Generation"
☆211Apr 1, 2025Updated last year
Alternatives and similar repositories for MakeAnything
Users that are interested in MakeAnything are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV 2025] Balanced Image Stylization with Style Matching Score☆69Mar 9, 2026Updated 4 months ago
- The official code implementation of the paper "OmniConsistency: Learning Style-Agnostic Consistency from Paired Stylization Data."☆423Jun 8, 2025Updated last year
- [CVPR 2026] An official implementation of Adv-GRPO. The Image as Its Own Reward: Reinforcement Learning with Adversarial Reward for Image…☆88Feb 26, 2026Updated 4 months ago
- Code for [CVPR 2025] ROICtrl: Boosting Instance Control for Visual Generation☆110Apr 16, 2025Updated last year
- A custom ComfyUI node designed to create seamless motion effects from single images by integrating with Hunyuan Video through latent spac…☆43Dec 27, 2024Updated last year
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Official code of "LayerTracer: Cognitive-Aligned Layered SVG Synthesis via Diffusion Transformer"☆98Apr 1, 2025Updated last year
- [NeurIPS 2024] EvolveDirector: Approaching Advanced Text-to-Image Generation with Large Vision-Language Models.☆52Oct 14, 2024Updated last year
- Node to tryoff clothes☆23Apr 14, 2025Updated last year
- [ICCV 2025] Code Implementation of "ArtEditor: Learning Customized Instructional Image Editor from Few-Shot Examples"☆430Apr 23, 2025Updated last year
- ICML 2025 - Impossible Videos☆81Jul 23, 2025Updated 11 months ago
- TPDiff: Temporal Pyramid Video Diffusion Model☆25Mar 13, 2025Updated last year
- [ICLR2025] A versatile image-to-image visual assistant, designed for image generation, manipulation, and translation based on free-from u…☆210May 5, 2025Updated last year
- [CVPR 2026] FocusUI: Efficient UI Grounding via Position-Preserving Visual Token Selection☆35Jun 7, 2026Updated last month
- [ICCV 2025] Official pytorch implementation of "FramePainter: Endowing Interactive Image Editing with Video Diffusion Priors"☆406Mar 10, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [NeurIPS 2025 Official Codes] Nabla-R2D3: Effective and Efficient 3D Diffusion Alignment with 2D Rewards☆45Sep 23, 2025Updated 9 months ago
- GenXD: Generating Any 3D and 4D Scenes. ICLR 2025☆225Mar 30, 2025Updated last year
- Training-free Regional Prompting for Diffusion Transformers 🔥☆696Nov 28, 2024Updated last year
- [SIGGRAPH 2025] Diffusion as Shader: 3D-aware Video Diffusion for Versatile Video Generation Control☆822Jun 9, 2025Updated last year
- [ICCV 2025 Highlight] OminiControl: Minimal and Universal Control for Diffusion Transformer☆1,925Jul 2, 2026Updated 2 weeks ago
- Piece it Together: Part-Based Concepting with IP-Priors☆93Apr 15, 2025Updated last year
- DMM: Building a Versatile Image Generation Model via Distillation-Based Model Merging☆47Apr 27, 2025Updated last year
- Official code for VMix: Improving Text-to-Image Diffusion Model with Cross-Attention Mixing Control☆191Dec 31, 2024Updated last year
- ☆246May 9, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Implementation code of the paper MIGE: A Unified Framework for Multimodal Instruction-Based Image Generation and Editing☆72Jul 13, 2025Updated last year
- Lets make loop video diffusion practical!☆85Jun 8, 2025Updated last year
- [NeurIPS 2025] IEAP: Image Editing As Programs with Diffusion Models☆118Sep 27, 2025Updated 9 months ago
- UniGeo: Taming Video Diffusion for Unified Consistent Geometry Estimation☆136Jun 10, 2025Updated last year
- [SIGGRAPH 2025] Official code of the paper "Cobra: Efficient Line Art COlorization with BRoAder References". Cobra:利用更广泛参考图实现高效线稿上色☆255Dec 10, 2025Updated 7 months ago
- [ICCV'25]DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion☆1,332Oct 17, 2025Updated 9 months ago
- UniWorld: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation☆883Dec 23, 2025Updated 6 months ago
- Code for: "Long-Context Autoregressive Video Modeling with Next-Frame Prediction"☆311Apr 23, 2025Updated last year
- [ICCV 2025] Light-A-Video: Training-free Video Relighting via Progressive Light Fusion☆517Oct 25, 2025Updated 8 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The code repository of UniRL☆53May 30, 2025Updated last year
- [CVPR 2025] Official repo for ART:Anonymous Region Transformer for Variable Multi-Layer Transparent Image Generation☆373Apr 8, 2026Updated 3 months ago
- [CVPR 2025 GMCV] Test-Time Frequency Scaling: Instant Frequency Control for Any Diffusion Model☆55May 31, 2025Updated last year
- [ICML 2026] Orienting Latent Actions for Video World Modeling☆116Apr 20, 2026Updated 3 months ago
- The official code of Refinement Provenance Inference: Detecting LLM-Refined Training Prompts from Model Behavior☆22Jan 6, 2026Updated 6 months ago
- [CVPR 2025 Highlight] X-Dyna: Expressive Dynamic Human Image Animation☆269Jan 30, 2025Updated last year
- Computer-Use Agents as Judges for Generative UI☆44Nov 27, 2025Updated 7 months ago