[ECCV 2024] Official pytorch implementation of "Switch Diffusion Transformer: Synergizing Denoising Tasks with Sparse Mixture-of-Experts"
☆48Jul 4, 2024Updated 2 years ago
Alternatives and similar repositories for Switch-DiT
Users that are interested in Switch-DiT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR 2024] Official pytorch implementation of "Denoising Task Routing for Diffusion Models"☆25Feb 19, 2024Updated 2 years ago
- [Neurips 2023] Official pytorch implementation of "Addressing Negative Transfer in Diffusion Models"☆25Jul 4, 2024Updated 2 years ago
- Google DeepMind: Mixture of Depths Unofficial Implementation.☆12May 29, 2024Updated 2 years ago
- [AAAI 2025] Official pytorch implementation of "Diffusion Model Patching via Mixture-of-Prompts"☆13Dec 12, 2024Updated last year
- Scaling Diffusion Transformers with Mixture of Experts☆433Sep 9, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A PyTorch implementation of the paper "ZigMa: A DiT-Style Mamba-based Diffusion Model" (ECCV 2024)☆351Mar 17, 2025Updated last year
- ☆16Sep 11, 2025Updated 10 months ago
- [ICCV 2025] Official pytorch implementation of "SteerX: Creating Any Camera-Free 3D and 4D Scenes with Geometric Steering"☆51Mar 20, 2025Updated last year
- [ICLR 2023] "Sparse MoE as the New Dropout: Scaling Dense and Self-Slimmable Transformers" by Tianlong Chen*, Zhenyu Zhang*, Ajay Jaiswal…☆56Feb 28, 2023Updated 3 years ago
- SpeeD: A Closer Look at Time Steps is Worthy of Triple Speed-Up for Diffusion Model Training☆188Jan 27, 2025Updated last year
- ☆20Nov 24, 2025Updated 8 months ago
- [NeurIPS 2024] RealCompo: Balancing Realism and Compositionality Improves Text-to-Image Diffusion Models☆121Nov 14, 2024Updated last year
- [CVPR '24] Official repository for Deformable One-shot Face Stylization via DINO Semantic Guidance☆60Apr 3, 2024Updated 2 years ago
- [Findings of ACL-2023] This is the official implementation of On the Difference of BERT-style and CLIP-style Text Encoders.☆14Jun 7, 2023Updated 3 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆24Jun 14, 2025Updated last year
- ☆25Dec 7, 2023Updated 2 years ago
- ☆56Apr 3, 2024Updated 2 years ago
- [ICML 2026] The official implementation of "Mean Flow Distillation: Robust and Stable Distillation for Flow Matching Models".☆23Jun 14, 2026Updated last month
- Official Implementation of VideoRFSplat: Direct Scene-Level Text-to-3D Gaussian Splatting Generation with Flexible Pose and Multi-View Jo…☆23Jun 27, 2025Updated last year
- The official implementation of "2025ICLR Dynamic Diffusion Transformer" and "2025ArXivDyDiT++: Dynamic Diffusion Transformers for Efficie…☆52Apr 10, 2025Updated last year
- [Arxiv 2025] Official PyTorch implementation of DiffMoE, TC-DiT, EC-DiT and Dense DiT☆175Oct 21, 2025Updated 9 months ago
- ☆18Aug 1, 2024Updated last year
- Code and benchmark for the paper: "A Practitioner's Guide to Continual Multimodal Pretraining" [NeurIPS'24]☆62Dec 10, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆18Jul 10, 2024Updated 2 years ago
- [IJCAI 2025] Offical implementation of the paper "Multi-View Learning with Context-Guided Receptance for Image Denoising".☆13Jun 26, 2025Updated last year
- Fast Kolmogorov-Arnold Network in JAX, initial experiments☆16May 20, 2024Updated 2 years ago
- An automatic MLLM hallucination detection framework☆19Sep 26, 2023Updated 2 years ago
- Code for Carve3D: Improving Multi-view Reconstruction Consistency for Diffusion Models with RL Finetuning☆37Jun 16, 2024Updated 2 years ago
- Official code for "Stitched Value Model for Diffusion Alignment"☆28May 21, 2026Updated 2 months ago
- Official implementation of "Perturbed-Attention Guidance"☆60Jul 2, 2024Updated 2 years ago
- ☆197Oct 28, 2024Updated last year
- Training-free Stylized Text-to-Image Generation with Fast Inference☆28May 30, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆60May 28, 2024Updated 2 years ago
- [CVPR 2026 (Highlight)] Unofficial Implementation of "Image Diffusion Preview with Consistency Solver"☆30Jan 24, 2026Updated 6 months ago
- Give us minutes, we give back a faster Mamba. The official implementation of "Faster Vision Mamba is Rebuilt in Minutes via Merged Token …☆39Dec 18, 2024Updated last year
- ☆20Sep 19, 2023Updated 2 years ago
- Semantic-Guided Diffusion Model for Single-Step Image Super-Resolution☆24Jun 10, 2025Updated last year
- [CVPR 2025] Multi-focal Conditioned Latent Diffusion for Person Image Synthesis☆23Mar 23, 2025Updated last year
- Face-MakeUp (SD1.5): Multimodal Facial Prompts for Text-to-Image Generation (ECAI-2025)☆26Jan 19, 2025Updated last year