Improves Text to Image synthesis from AttnGAN by integrating the scale-specific control from StyleGAN; can optionally use GPT-2 as text encoder
☆61Mar 22, 2022Updated 4 years ago
Alternatives and similar repositories for Style-AttnGAN
Users that are interested in Style-AttnGAN are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR2022 oral] A Simple and Effective Baseline for Text-to-Image Synthesis☆326Sep 24, 2025Updated 10 months ago
- Code for "Semantic Object Accuracy for Generative Text-to-Image Synthesis" (TPAMI 2020)☆106Jan 13, 2022Updated 4 years ago
- The official PyTorch implementation for MM'21 paper 'Attribute-specific Control Units in StyleGAN for Fine-grained Image Manipulation'☆39Dec 16, 2021Updated 4 years ago
- AttnGAN for the FashionGen Dataset☆71Jun 2, 2020Updated 6 years ago
- stylegan convert text to face image☆17Sep 29, 2019Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆14Aug 10, 2021Updated 5 years ago
- ☆194Aug 8, 2022Updated 4 years ago
- One Model to Edit Them All: Free-Form Text-Driven Image Manipulation with Semantic Modulations. NeurIPS2022.☆34Feb 13, 2023Updated 3 years ago
- Feed forward VQGAN-CLIP model, where the goal is to eliminate the need for optimizing the latent space of VQGAN for each input prompt☆140Jan 3, 2024Updated 2 years ago
- Official code repository for the EMNLP 2021 paper☆26Jan 30, 2022Updated 4 years ago
- ☆17Nov 4, 2022Updated 3 years ago
- [CVPR 2021] Multi-Modal-CelebA-HQ: A Large-Scale Text-Driven Face Generation and Understanding Dataset☆258Jun 1, 2024Updated 2 years ago
- Official code and data of "3AM: An Ambiguity-Aware Multi-Modal Machine Translation Dataset"☆12Dec 8, 2024Updated last year
- ☆1,363Jul 25, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for our Paper "One Model to Reconstruct Them All: A Novel Way to Use the Stochastic Noise in StyleGAN"☆73Nov 17, 2020Updated 5 years ago
- Text to Image Generation with Semantic-Spatial Aware GAN☆187Mar 24, 2022Updated 4 years ago
- A repository for the updated version of CoinRun used to collect MUGEN, a multimodal video-audio-text dataset. This repo contains scripts …☆13Jul 13, 2022Updated 4 years ago
- S2FGAN☆17Nov 7, 2021Updated 4 years ago
- ImageBART: Bidirectional Context with Multinomial Diffusion for Autoregressive Image Synthesis☆126Mar 14, 2022Updated 4 years ago
- Code Base for the work "Interactive Portrait Harmonization"☆28May 11, 2023Updated 3 years ago
- This repository is a repository for the paper, "Irgun: Improved residue based gradual up-scaling network for single image super resolutio…☆16Aug 26, 2020Updated 5 years ago
- Official code of "StyleT2I: Toward Compositional and High-Fidelity Text-to-Image Synthesis" (CVPR 2022)☆43Jul 31, 2022Updated 4 years ago
- ☆43Oct 16, 2021Updated 4 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Official repository for "Attend to Not Attended: Structure-then-Detail Token Merging for Post-training DiT Acceleration", which has been …☆17Sep 29, 2025Updated 10 months ago
- Commonality in Natural Images Rescues GANs: Pretraining GANs with Generic and Privacy-free Synthetic Data - Official PyTorch Implementati…☆34Nov 14, 2022Updated 3 years ago
- Password manager for shared accounts and device passwords, including LDAP integration.☆14Dec 16, 2014Updated 11 years ago
- [CVPR 2020] G3AN: Disentangling Appearance and Motion for Video Generation☆37Feb 5, 2021Updated 5 years ago
- Pytorch implementation for Controllable Text-to-Image Generation.☆169Dec 26, 2019Updated 6 years ago
- Adaptive Passage Encoder for Open-domain Question Answering☆15Jun 1, 2021Updated 5 years ago
- A modification of Daniel Russell's notebook merged with Katherine Crowson's hq-skip-net changes☆11Jan 28, 2022Updated 4 years ago
- ☆11Feb 24, 2022Updated 4 years ago
- VQA baseline with Conditional Batch Normalization☆15Apr 9, 2018Updated 8 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Lip and hair color editor using face parsing maps.☆11Jun 10, 2019Updated 7 years ago
- Official implementation of the paper Efficient Neural Architecture for Text-to-Image Synthesis.☆16Jun 8, 2022Updated 4 years ago
- Polysemous Visual-Semantic Embedding for Cross-Modal Retrieval (CVPR 2019)☆135Mar 15, 2024Updated 2 years ago
- Code for the paper "Modeling Information Change in Science Communication with Semantically Matched Paraphrases" from EMNLP 2022☆13Oct 20, 2022Updated 3 years ago
- ☆43Jul 18, 2022Updated 4 years ago
- ☆45Dec 26, 2021Updated 4 years ago
- Emofilt is a program to simulate emotional arousal with speech synthesis based on the free-for-non-commercial-use MBROLA synthesis engine…☆14Mar 17, 2022Updated 4 years ago