Predicting the generation FID of latent diffusion, with a variant of reconstruction FID of Variational Auto-encoder.
☆87Aug 28, 2026Updated 3 weeks ago
Alternatives and similar repositories for Making-rFID-Predictive-of-Diffusion-gFID
Users that are interested in Making-rFID-Predictive-of-Diffusion-gFID are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Metric implementation and raw data of "Diffusing in the Right Space: A Systematic Study of Latent Diffusability"☆42Jun 16, 2026Updated 3 months ago
- [ICCV 2025] Official implementation of the paper: REPA-E: Unlocking VAE for End-to-End Tuning of Latent Diffusion Transformers☆517Dec 6, 2025Updated 9 months ago
- [ECCV 2026] Towards Scalable Pre-training of Visual Tokenizers for Generation☆508Apr 15, 2026Updated 5 months ago
- Official code of Geometric Autoencoder for Diffusion Models.☆21Mar 12, 2026Updated 6 months ago
- (ICCV 2025) "Principal Components" Enable A New Language of Images☆89Jun 4, 2026Updated 3 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆569May 1, 2026Updated 4 months ago
- [ICCV2025] TokenBridge: Bridging Continuous and Discrete Tokens for Autoregressive Visual Generation. https://yuqingwang1029.github.io/To…☆159Jul 24, 2025Updated last year
- Official Implementation of "What Matters for Diffusion-Friendly Latent Manifold? Prior-Aligned Autoencoders for Latent Diffusion"☆83May 27, 2026Updated 3 months ago
- Unoffical Pytorch Implementation of Improving Inference for Neural Image Compression☆15Apr 27, 2025Updated last year
- Official PyTorch Implementation of "Diffusion Transformers with Representation Autoencoders"☆2,016Feb 25, 2026Updated 6 months ago
- Official implementation of (ICML 2026) Training-Free Vector Quantization via Gaussian VAEs☆26Aug 28, 2026Updated 3 weeks ago
- [CVPR 2025 Oral] Reconstruction vs. Generation: Taming Optimization Dilemma in Latent Diffusion Models☆1,540Dec 16, 2025Updated 9 months ago
- [ECCV 2026] Official code of "Representation Alignment for Just Image Transformers is not Easier than You Think"☆55Jun 18, 2026Updated 3 months ago
- Single-stage End-to-End Training for Tokenization and Generation☆121Mar 24, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official repository for “PixelGen: Improving Pixel Diffusion with Perceptual Loss”☆281May 12, 2026Updated 4 months ago
- [ICML26] Distribution Matching Variational AutoEncoder (DMVAE)☆58Dec 9, 2025Updated 9 months ago
- Towards Holistic evaluation of Generative Diffusion Transformers!☆104Jul 1, 2026Updated 2 months ago
- [ICLR 2026] Official implementation for What matters for Representation Alignment: Global Information or Spatial Structure?☆271Dec 15, 2025Updated 9 months ago
- Visual Generation Tuning☆102Apr 16, 2026Updated 5 months ago
- [ICML 2026] code & model for arxiv paper "Autoregressive Image Generation with Masked Bit Modeling"☆59May 1, 2026Updated 4 months ago
- CVPR 2026 (Highlight)-Guiding a Diffusion Transformer with the Internal Dynamics of Itself (IG)☆92Apr 9, 2026Updated 5 months ago
- PyTorch implementation of RiT: Vanilla Diffusion Transformers Suffice in Representation Space☆29May 23, 2026Updated 3 months ago
- Official PyTorch Implementation of "Latent Denoising Makes Good Visual Tokenizers"☆200Feb 24, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICLR2026] WeTok: Powerful Discrete Tokenization for High-Fidelity Visual Reconstruction☆71Sep 3, 2025Updated last year
- Official Implemenation for RAEv2: Improved Baselines with Representation Autoencoders☆326May 21, 2026Updated 3 months ago
- Implementation of the paper "Variable Bitrate Residual Vector Quantization for Audio Coding"☆11Apr 10, 2025Updated last year
- [CVPR2026 Highlight] Cubic Discrete Diffusion: Discrete Visual Generation on High-Dimensional Representation Tokens https://arxiv.org/abs…☆63Apr 10, 2026Updated 5 months ago
- [IEEE TCSVT] Preprocessing Enhanced Image Compression for Machine Vision☆19Mar 23, 2025Updated last year
- [CVPR2026 Highlight] Official repository for “DeCo: Frequency-Decoupled Pixel Diffusion for End-to-End Image Generation”☆243Feb 27, 2026Updated 6 months ago
- [ICML'26] Code and website for Self-Flow: Self-Supervised Flow Matching for Scalable Multi-Modal Synthesis☆766May 23, 2026Updated 3 months ago
- Official repo for UAE [ECCV 2026]☆213Jul 28, 2026Updated last month
- official implementation of the paper "Delving into Latent Spectral Biasing of Video VAEs for Superior Diffusability".☆83Dec 25, 2025Updated 8 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Official Implementation for the paper: A Variational Framework for Improving Naturalness in Generative Spoken Language Models☆24Jun 18, 2025Updated last year
- Frequency Autoregressive Image Generation with Continuous Tokens☆101Jun 9, 2025Updated last year
- FlexTok: Resampling Images into 1D Token Sequences of Flexible Length☆333Sep 11, 2026Updated last week
- Scaling Text-to-Image Diffusion Transformers with Representation Autoencoders☆263Feb 13, 2026Updated 7 months ago
- [AAAI 2026] Turbo-VAED: Fast and Stable Transfer of Video-VAEs to Mobile Devices☆142Jul 10, 2026Updated 2 months ago
- Transition Models☆156May 11, 2026Updated 4 months ago
- [ICLR'25 Oral] Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Think☆1,712Mar 16, 2025Updated last year