PiD: Fast and High-Resolution Latent Decoding with Pixel Diffusion
☆1,056Jul 22, 2026Updated last month
Alternatives and similar repositories for PiD
Users that are interested in PiD are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆565May 1, 2026Updated 4 months ago
- ☆109May 27, 2026Updated 3 months ago
- Official implementation of AsymFlow, pi-Flow, GMFlow☆465Jul 14, 2026Updated last month
- Flow Map OPD for AnyStep Video Diffusion☆421Aug 14, 2026Updated 3 weeks ago
- Official implementation of Tuna-2: Pixel Embeddings Beat Vision Encoders for Unified Understanding and Generation☆755Jul 22, 2026Updated last month
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [ICML'26] Code and website for Self-Flow: Self-Supervised Flow Matching for Scalable Multi-Modal Synthesis☆765May 23, 2026Updated 3 months ago
- [CVPR 2026 Best Paper Finalist] Pixel Diffusion Transformers for Image Generation☆948Jul 8, 2026Updated 2 months ago
- Code for "L2P: Unlocking Latent Potential for Pixel Generation"☆191Jul 11, 2026Updated last month
- Official repository for “PixelGen: Improving Pixel Diffusion with Perceptual Loss”☆278May 12, 2026Updated 3 months ago
- rCM & Causal-rCM: Leading and Unified Algorithms/Infrastructures for Bidirectional/Autoregressive Video Diffusion Distillation at Scale☆800Jun 25, 2026Updated 2 months ago
- Official implementation of "SEGA: Spectral-Energy Guided Attention for Resolution Extrapolation in Diffusion Transformers"☆73Jul 4, 2026Updated 2 months ago
- [ICLR 2026] ChronoEdit: Towards Temporal Reasoning for Image Editing and World Simulation☆706Nov 20, 2025Updated 9 months ago
- Official Implemenation for RAEv2: Improved Baselines with Representation Autoencoders☆324May 21, 2026Updated 3 months ago
- Boogu-Image-0.1 is an Apache-2.0 open-source image generation and editing model family that delivers near-closed-source performance with …☆981Jul 23, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICLR 2026 Oral] DiffusionNFT: Online Diffusion Reinforcement with Forward Process☆1,050Feb 10, 2026Updated 6 months ago
- [ICML 2026] Official codebase for "Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactiv…☆951Aug 28, 2026Updated last week
- Helios: Real Real-Time Long Video Generation Model☆2,136Aug 24, 2026Updated 2 weeks ago
- Scaling Text-to-Image Diffusion Transformers with Representation Autoencoders☆263Feb 13, 2026Updated 6 months ago
- [ICLR 2026] UniVideo: Unified Understanding, Generation, and Editing for Videos☆556Jul 3, 2026Updated 2 months ago
- A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.☆1,335Jul 14, 2026Updated last month
- Official Repo of "D-OPSD: On-Policy Self-Distillation for Continuously Tuning Step-Distilled Diffusion Models"☆324May 22, 2026Updated 3 months ago
- [ICLR 2026] Taming large-scale few-step training with self-adversarial flows! 👏🏻☆538Feb 24, 2026Updated 6 months ago
- Code release for "i1: A Simple and Fully Open Recipe for Strong Text-to-Image Models"☆287Aug 16, 2026Updated 3 weeks ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Continuous-Time Distribution Matching for Few-Step Diffusion Distillation👏☆153May 11, 2026Updated 3 months ago
- DreamX-World: A General-Purpose Interactive World Model☆767Jul 23, 2026Updated last month
- HY-World 1.5: A Systematic Framework for Interactive World Modeling with Real-Time Latency and Geometric Consistency☆1,600Jun 10, 2026Updated 2 months ago
- ☆84Jul 3, 2026Updated 2 months ago
- [ICLR 2026] Official Repo for Rolling Forcing: Autoregressive Long Video Diffusion in Real Time☆459Oct 31, 2025Updated 10 months ago
- [NeurIPS 2025 Oral]Infinity⭐️: Unified Spacetime AutoRegressive Modeling for Visual Generation☆783Apr 16, 2026Updated 4 months ago
- Warp-as-History: Generalizable Camera-Controlled Video Generation from One Training Video☆226May 30, 2026Updated 3 months ago
- Official code, models, and data for Vista4D: Video Reshooting with 4D Point Clouds (CVPR 2026 Highlight)☆588Jun 2, 2026Updated 3 months ago
- Cameras as Relative Positional Encoding☆749Dec 18, 2025Updated 8 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- UltraFlux: Data-Model Co-Design for High-quality Native 4K Text-to-Image Generation across Diverse Aspect Ratios☆149Apr 9, 2026Updated 4 months ago
- Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence☆946Aug 5, 2026Updated last month
- Project Lyra: Open Generative 3D World Models☆2,292Jul 20, 2026Updated last month
- A Minimal and Elegant Framework & Tutorial for Real-Time Interactive World Models☆812Updated this week
- (NeurIPS 2024 Oral 🔥) Improved Distribution Matching Distillation for Fast Image Synthesis☆1,450Mar 5, 2025Updated last year
- Bernini is a unified framework for video generation and editing that combines an MLLM-based semantic planner with a DiT-based renderer.☆1,308Aug 13, 2026Updated 3 weeks ago
- NVIDIA FastGen: Fast Generation from Diffusion Models☆990Aug 21, 2026Updated 2 weeks ago