Predicting the generation FID of latent diffusion, with a variant of reconstruction FID of Variational Auto-encoder.
☆88Aug 28, 2026Updated this week
Alternatives and similar repositories for Making-rFID-Predictive-of-Diffusion-gFID
Users that are interested in Making-rFID-Predictive-of-Diffusion-gFID are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Metric implementation and raw data of "Diffusing in the Right Space: A Systematic Study of Latent Diffusability"☆42Jun 16, 2026Updated 2 months ago
- [ICCV 2025] Official implementation of the paper: REPA-E: Unlocking VAE for End-to-End Tuning of Latent Diffusion Transformers☆514Dec 6, 2025Updated 8 months ago
- [ECCV 2026] Towards Scalable Pre-training of Visual Tokenizers for Generation☆504Apr 15, 2026Updated 4 months ago
- Official code of Geometric Autoencoder for Diffusion Models.☆21Mar 12, 2026Updated 5 months ago
- (ICCV 2025) "Principal Components" Enable A New Language of Images☆88Jun 4, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆560May 1, 2026Updated 3 months ago
- [ICCV2025] TokenBridge: Bridging Continuous and Discrete Tokens for Autoregressive Visual Generation. https://yuqingwang1029.github.io/To…☆158Jul 24, 2025Updated last year
- Official Implementation of "What Matters for Diffusion-Friendly Latent Manifold? Prior-Aligned Autoencoders for Latent Diffusion"☆81May 27, 2026Updated 3 months ago
- Unoffical Pytorch Implementation of Improving Inference for Neural Image Compression☆15Apr 27, 2025Updated last year
- Official PyTorch Implementation of "Diffusion Transformers with Representation Autoencoders"☆2,001Feb 25, 2026Updated 6 months ago
- Official implementation of (ICML 2026) Training-Free Vector Quantization via Gaussian VAEs☆26Updated this week
- [CVPR 2025 Oral] Reconstruction vs. Generation: Taming Optimization Dilemma in Latent Diffusion Models☆1,533Dec 16, 2025Updated 8 months ago
- [ECCV 2026] Official code of "Representation Alignment for Just Image Transformers is not Easier than You Think"☆54Jun 18, 2026Updated 2 months ago
- Single-stage End-to-End Training for Tokenization and Generation☆121Mar 24, 2026Updated 5 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Official repository for “PixelGen: Improving Pixel Diffusion with Perceptual Loss”☆277May 12, 2026Updated 3 months ago
- [ICML26] Distribution Matching Variational AutoEncoder (DMVAE)☆57Dec 9, 2025Updated 8 months ago
- Towards Holistic evaluation of Generative Diffusion Transformers!☆104Jul 1, 2026Updated last month
- [ICLR 2026] Official implementation for What matters for Representation Alignment: Global Information or Spatial Structure?☆269Dec 15, 2025Updated 8 months ago
- Visual Generation Tuning☆101Apr 16, 2026Updated 4 months ago
- [ICML 2026] code & model for arxiv paper "Autoregressive Image Generation with Masked Bit Modeling"☆60May 1, 2026Updated 3 months ago
- CVPR 2026 (Highlight)-Guiding a Diffusion Transformer with the Internal Dynamics of Itself (IG)☆90Apr 9, 2026Updated 4 months ago
- PyTorch implementation of RiT: Vanilla Diffusion Transformers Suffice in Representation Space☆28May 23, 2026Updated 3 months ago
- Official PyTorch Implementation of "Latent Denoising Makes Good Visual Tokenizers"☆197Feb 24, 2026Updated 6 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [ICLR2026] WeTok: Powerful Discrete Tokenization for High-Fidelity Visual Reconstruction☆71Sep 3, 2025Updated 11 months ago
- Official Implemenation for RAEv2: Improved Baselines with Representation Autoencoders☆324May 21, 2026Updated 3 months ago
- Implementation of the paper "Variable Bitrate Residual Vector Quantization for Audio Coding"☆11Apr 10, 2025Updated last year
- [CVPR2026 Highlight] Cubic Discrete Diffusion: Discrete Visual Generation on High-Dimensional Representation Tokens https://arxiv.org/abs…☆63Apr 10, 2026Updated 4 months ago
- [IEEE TCSVT] Preprocessing Enhanced Image Compression for Machine Vision☆19Mar 23, 2025Updated last year
- [CVPR2026 Highlight] Official repository for “DeCo: Frequency-Decoupled Pixel Diffusion for End-to-End Image Generation”☆240Feb 27, 2026Updated 6 months ago
- [ICML'26] Code and website for Self-Flow: Self-Supervised Flow Matching for Scalable Multi-Modal Synthesis☆760May 23, 2026Updated 3 months ago
- official implementation of the paper "Delving into Latent Spectral Biasing of Video VAEs for Superior Diffusability".☆78Dec 25, 2025Updated 8 months ago
- Frequency Autoregressive Image Generation with Continuous Tokens☆101Jun 9, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- FlexTok: Resampling Images into 1D Token Sequences of Flexible Length☆331Jun 2, 2025Updated last year
- Scaling Text-to-Image Diffusion Transformers with Representation Autoencoders☆262Feb 13, 2026Updated 6 months ago
- [AAAI 2026] Turbo-VAED: Fast and Stable Transfer of Video-VAEs to Mobile Devices☆140Jul 10, 2026Updated last month
- Official repo for UAE [ECCV 2026]☆211Jul 28, 2026Updated last month
- Transition Models☆156May 11, 2026Updated 3 months ago
- [ICLR'25 Oral] Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Think☆1,705Mar 16, 2025Updated last year
- PyTorch implementation of JiT https://arxiv.org/abs/2511.13720☆2,509Dec 8, 2025Updated 8 months ago