An in-context conditioning version of MUSE with pre-trained checkpoints.
☆115Jun 4, 2023Updated 3 years ago
Alternatives and similar repositories for MUSE-Pytorch
Users that are interested in MUSE-Pytorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Open reproduction of MUSE for fast text2image generation.☆359Jun 1, 2024Updated 2 years ago
- Unoffical implement for [StyleDrop](https://arxiv.org/abs/2306.00983)☆588Aug 23, 2023Updated 2 years ago
- This is an unofficial PyTorch implementation of StyleDrop: Text-to-Image Generation in Any Style.☆226Jul 11, 2023Updated 3 years ago
- [CVPR 2024] CapsFusion: Rethinking Image-Text Data at Scale☆215Feb 27, 2024Updated 2 years ago
- Paper List for In-context Learning 🌷☆19Jan 3, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆89Jan 4, 2024Updated 2 years ago
- Implementation of Muse: Text-to-Image Generation via Masked Generative Transformers, in Pytorch☆918Feb 29, 2024Updated 2 years ago
- A PyTorch implementation of the paper "All are Worth Words: A ViT Backbone for Diffusion Models".☆1,108Mar 25, 2023Updated 3 years ago
- recipe for training fully-featured self supervised image jepa models☆14Jun 4, 2025Updated last year
- Obj2Seq: Formatting Objects as Sequences with Class Prompt for Visual Tasks (NeurIPS2022)☆85Nov 2, 2022Updated 3 years ago
- Official PyTorch implementation of the paper "In-Context Learning Unlocked for Diffusion Models"☆414Mar 25, 2024Updated 2 years ago
- [NeurIPS 2022] code for "Visual Concepts Tokenization"☆23Oct 10, 2022Updated 3 years ago
- Zero-Shot Video Editing Using Off-The-Shelf Image Diffusion Models☆356Jul 4, 2023Updated 3 years ago
- ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation (ICCV 2023, Oral)☆541Jan 8, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code and models for the paper "One Transformer Fits All Distributions in Multi-Modal Diffusion"☆1,484May 31, 2023Updated 3 years ago
- unofficial pytorch implementation of HiFi-GAN with fast MISR.☆15Mar 21, 2023Updated 3 years ago
- Test-Time Training on Video Streams☆70Jul 24, 2023Updated 3 years ago
- API to extract data from wikiHow☆18Jul 10, 2021Updated 5 years ago
- Official PyTorch Implementation of "SiT: Exploring Flow and Diffusion-based Generative Models with Scalable Interpolant Transformers"☆1,200Dec 22, 2025Updated 7 months ago
- [CVPR'23] Video Probabilistic Diffusion Models in Projected Latent Space☆322May 14, 2024Updated 2 years ago
- ☆64Jul 1, 2023Updated 3 years ago
- Code release for Deep Incubation (https://arxiv.org/abs/2212.04129)☆92Mar 16, 2023Updated 3 years ago
- Official Jax Implementation of MaskGIT☆562Nov 18, 2022Updated 3 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Versatile Diffusion: Text, Images and Variations All in One Diffusion Model, arXiv 2022 / ICCV 2023☆1,334Aug 10, 2023Updated 3 years ago
- A PyTorch implementation of MAGE: MAsked Generative Encoder to Unify Representation Learning and Image Synthesis☆581Mar 10, 2023Updated 3 years ago
- Exploiting unlabeled data with vision and language models for object detection, ECCV 2022☆96Jan 16, 2024Updated 2 years ago
- This repo contains the code for 1D tokenizer and generator☆1,172Mar 20, 2025Updated last year
- Official PyTorch Implementation for Shape-Guided Diffusion with Inside-Outside Attention, WACV 2024☆39Aug 19, 2023Updated 2 years ago
- Phonemes and durations labeling based on whisper small☆11Jul 7, 2024Updated 2 years ago
- UEval: A Benchmark for Unified Multimodal Generation☆25Apr 20, 2026Updated 3 months ago
- [NeurIPS 2023] Text data, code and pre-trained models for paper "Improving CLIP Training with Language Rewrites"☆291Jan 14, 2024Updated 2 years ago
- [IJCV'24] AutoStory: Generating Diverse Storytelling Images with Minimal Human Effort☆149Mar 5, 2026Updated 5 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- code release of research paper "Exploring Long-Sequence Masked Autoencoders"☆100Oct 14, 2022Updated 3 years ago
- [CVPR 2023] RILS: Masked Visual Reconstruction in Language Semantic Space (https://arxiv.org/abs/2301.06958)☆44Sep 5, 2023Updated 2 years ago
- MoVQGAN - model for the image encoding and reconstruction☆266Oct 31, 2023Updated 2 years ago
- Mini-DALLE3: Interactive Text to Image by Prompting Large Language Models☆313Dec 28, 2023Updated 2 years ago
- 🤗 Unofficial huggingface/diffusers-based implementation of the paper "Training-Free Structured Diffusion Guidance for Compositional Text…☆120Mar 29, 2023Updated 3 years ago
- Emu Series: Generative Multimodal Models from BAAI☆1,777Jan 12, 2026Updated 7 months ago
- ☆197Dec 17, 2024Updated last year