A Gradio demo of MGIE
☆346Feb 23, 2024Updated 2 years ago
Alternatives and similar repositories for pytorch_mgie
Users that are interested in pytorch_mgie are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆3,874Mar 15, 2024Updated 2 years ago
- ☆25May 12, 2026Updated 2 months ago
- ☆13Mar 15, 2024Updated 2 years ago
- ☆13Feb 28, 2024Updated 2 years ago
- ☆14Dec 26, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Python scripts to create noisy and reverberant 2-speaker mixture audio with Libri-Light and WHAM☆17Nov 7, 2024Updated last year
- ☆25Dec 21, 2023Updated 2 years ago
- ☆17Jan 2, 2024Updated 2 years ago
- we propose FlexEdit, an end-to-end image editing method that leverages both free-shape masks and language instructions for Flexible Editi…☆32Aug 22, 2024Updated last year
- ☆24Dec 31, 2023Updated 2 years ago
- Official code of SmartEdit [CVPR-2024 Highlight]☆374Jun 21, 2024Updated 2 years ago
- 🐟 Code and models for the NeurIPS 2023 paper "Generating Images with Multimodal Language Models".☆470Jan 19, 2024Updated 2 years ago
- Official code for the CVPR 2025 paper "SemanticDraw: Towards Real-Time Interactive Content Creation from Image Diffusion Models."☆589Jun 1, 2025Updated last year
- UnifiedMLLM: Enabling Unified Representation for Multi-modal Multi-tasks With Large Language Model☆22Aug 5, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆17Dec 18, 2023Updated 2 years ago
- Multimodal Models in Real World☆558Feb 24, 2025Updated last year
- PyTorch implementation of InstructDiffusion, a unifying and generic framework for aligning computer vision tasks with human instructions.☆445May 14, 2024Updated 2 years ago
- [ACM MM 2024] Offical Code for "HandRefiner: Refining Malformed Hands in Generated Images by Diffusion-based Conditional Inpainting"☆806Oct 31, 2024Updated last year
- ☆24Dec 10, 2023Updated 2 years ago
- [ACM Multimedia 2025 Datasets Track] EditWorld: Simulating World Dynamics for Instruction-Following Image Editing☆141Aug 2, 2025Updated 11 months ago
- [ECCV 2024] FreeInit: Bridging Initialization Gap in Video Diffusion Models☆544Jan 18, 2024Updated 2 years ago
- ☆18Dec 29, 2023Updated 2 years ago
- ☆77Dec 25, 2023Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [CVPR 2024] U-VAP: User-specified Visual Appearance Personalization via Decoupled Self Augmentation☆18Sep 1, 2024Updated last year
- ☆15Sep 8, 2023Updated 2 years ago
- [ICLR 2025] HQ-Edit: A High-Quality and High-Coverage Dataset for General Image Editing☆114Apr 18, 2024Updated 2 years ago
- ☆17Dec 28, 2023Updated 2 years ago
- Implementation of "SCEdit: Efficient and Controllable Image Diffusion Generation via Skip Connection Editing"☆85Dec 26, 2023Updated 2 years ago
- [ICML 2024] Mastering Text-to-Image Diffusion: Recaptioning, Planning, and Generating with Multimodal LLMs (RPG)☆1,838Feb 1, 2025Updated last year
- Animated optical illusions in ComfyUI☆21Jun 14, 2024Updated 2 years ago
- 🏞️ Official implementation of "Gen4Gen: Generative Data Pipeline for Generative Multi-Concept Composition"☆110Mar 27, 2026Updated 3 months ago
- ☆18Dec 18, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- PhotoMaker [CVPR 2024]☆10,099Oct 31, 2024Updated last year
- The implementation of "Learning by Planning: Language-Guided Global Image Editing"☆25May 10, 2023Updated 3 years ago
- ☆411Jun 14, 2024Updated 2 years ago
- SCEPTER is an open-source framework used for training, fine-tuning, and inference with generative models.☆549Apr 3, 2025Updated last year
- Official implementation for "pOps: Photo-Inspired Diffusion Operators"☆86Jul 23, 2024Updated last year
- Repository for the Paper "Multi-LoRA Composition for Image Generation"☆492Mar 31, 2024Updated 2 years ago
- ViT models pretrained with up to ~5k hours of human-like video data☆14Aug 10, 2023Updated 2 years ago