An in-context conditioning version of MUSE with pre-trained checkpoints.
☆115Jun 4, 2023Updated 3 years ago
Alternatives and similar repositories for MUSE-Pytorch
Users that are interested in MUSE-Pytorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Unoffical implement for [StyleDrop](https://arxiv.org/abs/2306.00983)☆588Aug 23, 2023Updated 2 years ago
- This is an unofficial PyTorch implementation of StyleDrop: Text-to-Image Generation in Any Style.☆226Jul 11, 2023Updated 3 years ago
- [CVPR 2024] CapsFusion: Rethinking Image-Text Data at Scale☆215Feb 27, 2024Updated 2 years ago
- Paper List for In-context Learning 🌷☆19Jan 3, 2023Updated 3 years ago
- ☆89Jan 4, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Implementation of Muse: Text-to-Image Generation via Masked Generative Transformers, in Pytorch☆918Feb 29, 2024Updated 2 years ago
- A PyTorch implementation of the paper "All are Worth Words: A ViT Backbone for Diffusion Models".☆1,107Mar 25, 2023Updated 3 years ago
- recipe for training fully-featured self supervised image jepa models☆14Jun 4, 2025Updated last year
- Obj2Seq: Formatting Objects as Sequences with Class Prompt for Visual Tasks (NeurIPS2022)☆85Nov 2, 2022Updated 3 years ago
- Official PyTorch implementation of the paper "In-Context Learning Unlocked for Diffusion Models"☆414Mar 25, 2024Updated 2 years ago
- [NeurIPS 2022] code for "Visual Concepts Tokenization"☆23Oct 10, 2022Updated 3 years ago
- Zero-Shot Video Editing Using Off-The-Shelf Image Diffusion Models☆356Jul 4, 2023Updated 3 years ago
- ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation (ICCV 2023, Oral)☆541Jan 8, 2024Updated 2 years ago
- Code and models for the paper "One Transformer Fits All Distributions in Multi-Modal Diffusion"☆1,486May 31, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- unofficial pytorch implementation of HiFi-GAN with fast MISR.☆15Mar 21, 2023Updated 3 years ago
- Test-Time Training on Video Streams☆70Jul 24, 2023Updated 2 years ago
- API to extract data from wikiHow☆18Jul 10, 2021Updated 5 years ago
- Official PyTorch Implementation of "SiT: Exploring Flow and Diffusion-based Generative Models with Scalable Interpolant Transformers"☆1,188Dec 22, 2025Updated 7 months ago
- [CVPR'23] Video Probabilistic Diffusion Models in Projected Latent Space☆322May 14, 2024Updated 2 years ago
- ☆64Jul 1, 2023Updated 3 years ago
- Code release for Deep Incubation (https://arxiv.org/abs/2212.04129)☆92Mar 16, 2023Updated 3 years ago
- Official Jax Implementation of MaskGIT☆562Nov 18, 2022Updated 3 years ago
- Versatile Diffusion: Text, Images and Variations All in One Diffusion Model, arXiv 2022 / ICCV 2023☆1,334Aug 10, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A PyTorch implementation of MAGE: MAsked Generative Encoder to Unify Representation Learning and Image Synthesis☆582Mar 10, 2023Updated 3 years ago
- Exploiting unlabeled data with vision and language models for object detection, ECCV 2022☆97Jan 16, 2024Updated 2 years ago
- This repo contains the code for 1D tokenizer and generator☆1,166Mar 20, 2025Updated last year
- Official PyTorch Implementation for Shape-Guided Diffusion with Inside-Outside Attention, WACV 2024☆39Aug 19, 2023Updated 2 years ago
- Phonemes and durations labeling based on whisper small☆11Jul 7, 2024Updated 2 years ago
- UEval: A Benchmark for Unified Multimodal Generation☆24Apr 20, 2026Updated 3 months ago
- [NeurIPS 2023] Text data, code and pre-trained models for paper "Improving CLIP Training with Language Rewrites"☆291Jan 14, 2024Updated 2 years ago
- [IJCV'24] AutoStory: Generating Diverse Storytelling Images with Minimal Human Effort☆149Mar 5, 2026Updated 4 months ago
- code release of research paper "Exploring Long-Sequence Masked Autoencoders"☆100Oct 14, 2022Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [CVPR 2023] RILS: Masked Visual Reconstruction in Language Semantic Space (https://arxiv.org/abs/2301.06958)☆44Sep 5, 2023Updated 2 years ago
- MoVQGAN - model for the image encoding and reconstruction☆266Oct 31, 2023Updated 2 years ago
- Mini-DALLE3: Interactive Text to Image by Prompting Large Language Models☆313Dec 28, 2023Updated 2 years ago
- 🤗 Unofficial huggingface/diffusers-based implementation of the paper "Training-Free Structured Diffusion Guidance for Compositional Text…☆120Mar 29, 2023Updated 3 years ago
- Emu Series: Generative Multimodal Models from BAAI☆1,776Jan 12, 2026Updated 6 months ago
- ☆196Dec 17, 2024Updated last year
- TVMScript kernel for deformable attention☆25Dec 15, 2021Updated 4 years ago