Pytorch implementation of Make-A-Scene: Scene-Based Text-to-Image Generation with Human Priors
☆336Aug 9, 2022Updated 4 years ago
Alternatives and similar repositories for Make-A-Scene
Users that are interested in Make-A-Scene are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation for "Blended Diffusion for Text-driven Editing of Natural Images" [CVPR 2022]☆588Jun 4, 2024Updated 2 years ago
- ☆242Apr 5, 2022Updated 4 years ago
- [SIGGRAPH'22] StyleGAN-XL: Scaling StyleGAN to Large Diverse Datasets☆995Jun 24, 2024Updated 2 years ago
- Implementation of NÜWA, state of the art attention network for text to video synthesis, in Pytorch☆547Jan 17, 2023Updated 3 years ago
- Official implementation of VQ-Diffusion☆985Apr 17, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆194Dec 7, 2021Updated 4 years ago
- Styled text-to-drawing synthesis method. Featured at IJCAI 2022 and the 2021 NeurIPS Workshop on Machine Learning for Creativity and Desi…☆280Nov 15, 2022Updated 3 years ago
- Implementation of Parti, Google's pure attention-based text-to-image neural network, in Pytorch☆536Dec 8, 2023Updated 2 years ago
- ☆1,195Sep 29, 2022Updated 3 years ago
- Majesty Diffusion by @Dango233(@Dango233max) and @apolinario (@multimodalart)☆275Jul 25, 2022Updated 4 years ago
- 1.4B latent diffusion model fine tuning☆266May 16, 2022Updated 4 years ago
- Code for paper LAFITE: Towards Language-Free Training for Text-to-Image Generation (CVPR 2022)☆185Mar 23, 2023Updated 3 years ago
- ☆487Jun 30, 2022Updated 4 years ago
- official code repo for paper "CogView2: Faster and Better Text-to-Image Generation via Hierarchical Transformers"☆951Aug 3, 2022Updated 4 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- v objective diffusion inference code for PyTorch.☆720Nov 29, 2022Updated 3 years ago
- [ECCV 2022] Compositional Generation using Diffusion Models☆490Apr 24, 2025Updated last year
- Unified API to facilitate usage of pre-trained "perceptor" models, a la CLIP☆39Nov 26, 2022Updated 3 years ago
- combination of OpenAI GLIDE and Latent Diffusion☆135Apr 7, 2022Updated 4 years ago
- ☆3,055Feb 27, 2023Updated 3 years ago
- GLIDE: a diffusion-based text-conditional image synthesis model☆3,678Mar 8, 2024Updated 2 years ago
- A notebook for text-based guided image generation using StyleGANXL and CLIP.☆58May 19, 2023Updated 3 years ago
- ☆1,590Jun 28, 2022Updated 4 years ago
- Official Implementation for "HyperStyle: StyleGAN Inversion with HyperNetworks for Real Image Editing" (CVPR 2022) https://arxiv.org/abs/…☆1,028Sep 17, 2022Updated 4 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- PITI: Pretraining is All You Need for Image-to-Image Translation☆503Jun 2, 2024Updated 2 years ago
- Text-to-Image generation. The repo for NeurIPS 2021 paper "CogView: Mastering Text-to-Image Generation via Transformers".☆1,799Sep 25, 2023Updated 3 years ago
- A collection of generative and training notebooks getting mirrored to google colab.☆12May 29, 2022Updated 4 years ago
- CLOOB Conditioned Latent Diffusion training and inference code☆112Apr 15, 2022Updated 4 years ago
- Versatile Diffusion: Text, Images and Variations All in One Diffusion Model, arXiv 2022 / ICCV 2023☆1,332Aug 10, 2023Updated 3 years ago
- Official Pytorch Implementation for "Text2LIVE: Text-Driven Layered Image and Video Editing" (ECCV 2022 Oral)☆885Mar 9, 2023Updated 3 years ago
- A CLI tool/python module for generating images from text using guided diffusion and CLIP from OpenAI.☆459Dec 31, 2025Updated 8 months ago
- Feed forward VQGAN-CLIP model, where the goal is to eliminate the need for optimizing the latent space of VQGAN for each input prompt☆140Jan 3, 2024Updated 2 years ago
- [CVPR 2022] Official PyTorch Implementation for DiffusionCLIP: Text-guided Image Manipulation Using Diffusion Models☆864Mar 27, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Official Implementation of Paella https://arxiv.org/abs/2211.07292v2☆747Oct 4, 2023Updated 2 years ago
- ☆3,456May 14, 2024Updated 2 years ago
- Tackling the Generative Learning Trilemma with Denoising Diffusion GANs https://arxiv.org/abs/2112.07804☆760Dec 2, 2022Updated 3 years ago
- Home of `erlich` and `ongo`. Finetune latent-diffusion/glid-3-xl text2image on your own data.☆182Aug 5, 2022Updated 4 years ago
- The code of '3D-Aware Semantic-Guided Generative Model for Human Synthesis' (ECCV 2022)☆36Jul 18, 2022Updated 4 years ago
- [CVPR 2022] GAN inversion and editing with spatially-adaptive multiple latent layers☆173Jan 21, 2023Updated 3 years ago
- ☆152Sep 28, 2022Updated 3 years ago