The official implementation of Latte: Latent Diffusion Transformer for Video Generation.
☆34Feb 26, 2025Updated last year
Alternatives and similar repositories for Latte
Users that are interested in Latte are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV 2025] QuEST: Efficient Finetuning for Low-bit Diffusion Models☆61Jun 26, 2025Updated last year
- PyTorch implementation of PTQ4DiT https://arxiv.org/abs/2405.16005☆49Nov 8, 2024Updated last year
- Reuse and Diffuse: Iterative Denoising for Text-to-Video Generation☆38Nov 21, 2023Updated 2 years ago
- ☆17Jan 10, 2024Updated 2 years ago
- ☆11Jan 16, 2024Updated 2 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- [TMLR 2025] Latte: Latent Diffusion Transformer for Video Generation.☆1,948Aug 10, 2026Updated 3 weeks ago
- The code for the Network Binarization via Contrastive Learning, which has been accepted to ECCV 2022.☆14Jul 13, 2022Updated 4 years ago
- Activation-aware Singular Value Decomposition for Compressing Large Language Models☆92Oct 22, 2024Updated last year
- Implementation of Post-training Quantization on Diffusion Models (CVPR 2023)☆146Apr 1, 2023Updated 3 years ago
- [ICCV 2023 Oral, Best Paper Finalist] ITI-GEN: Inclusive Text-to-Image Generation☆67Feb 16, 2024Updated 2 years ago
- Super simple Streamlit app for playing with Stable Diffusion 2 and Stable Diffusion XL 1.0☆27Aug 23, 2024Updated 2 years ago
- 🧂 [ECCV 2026] Salt: Self-Consistent Distribution Matching with Cache-Aware Training for Fast Video Generation☆20Aug 24, 2026Updated last week
- 👆Pytorch implementation of "Ctrl-V: Higher Fidelity Video Generation with Bounding-Box Controlled Object Motion"☆36Jul 28, 2025Updated last year
- Official Implementation of paper "Distilling Long-tailed Datasets" [CVPR 2025]☆25Aug 13, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- [ICLR 2024 Spotlight] This is the official PyTorch implementation of "EfficientDM: Efficient Quantization-Aware Fine-Tuning of Low-Bit Di…☆74Jun 4, 2024Updated 2 years ago
- The HD-VG-130M Dataset☆126Apr 8, 2024Updated 2 years ago
- ☆16Apr 30, 2026Updated 4 months ago
- [ICML 2025] This is the official PyTorch implementation of "ZipAR: Accelerating Auto-regressive Image Generation through Spatial Locality…☆52Mar 25, 2025Updated last year
- ☆13Aug 18, 2026Updated 2 weeks ago
- Animated visualizers for signals.☆12Jun 23, 2026Updated 2 months ago
- [Preprint 2025] ICVE: In-Context Learning with Unpaired Clips for Instruction-based Video Editing☆26Jun 2, 2026Updated 2 months ago
- ☆16Dec 1, 2025Updated 9 months ago
- [ICCV 2023] PATMAT Person Aware Tuning of Mask Aware Transformer for Face Inpainting☆29Jan 5, 2024Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- FA4-based Relative Attention Kernel developed by TML and Colfax☆18Jul 17, 2026Updated last month
- [ACM MM24] MotionMaster: Training-free Camera Motion Transfer For Video Generation☆103Oct 15, 2024Updated last year
- Methodological considerations for analyzing and interpreting neural oscillations.☆31Oct 29, 2023Updated 2 years ago
- Streaming Video Diffusion: Online Video Editing with Diffusion Models☆17Jun 3, 2024Updated 2 years ago
- Sci-Fi: Symmetric Constraint for Frame Inbetweening☆21Aug 12, 2025Updated last year
- [AAAI-24] VVS : Video-to-Video Retrieval With Irrelevant Frame Suppression☆21May 14, 2024Updated 2 years ago
- [AAAI 2025] Official Implementation for "Click2Mask: Local Editing with Dynamic Mask Generation" Paper.☆21Jan 22, 2026Updated 7 months ago
- ☆18Sep 7, 2023Updated 2 years ago
- ☆26Sep 15, 2022Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- The codes of paper: "Reconstructing Perceptive Images from Brain Activity by Shape-Semantic GAN". This demo takes the published fMRI data…☆13Dec 12, 2020Updated 5 years ago
- Code for Semantically Robust Unpaired Image Translation for Data with Unmatched Semantics Statistics (SRUNIT), ICCV 2021☆11Feb 10, 2022Updated 4 years ago
- Score identity Distillation with Long and Short Guidance for One-Step Text-to-Image Generation☆96Dec 4, 2025Updated 8 months ago
- Official code base of "Perception-Oriented Video Frame Interpolation via Asymmetric Blending" (CVPR 2024), also denoted as ''PerVFI''.☆81Jan 11, 2025Updated last year
- ☆12Jul 3, 2023Updated 3 years ago
- ☆14Jul 22, 2021Updated 5 years ago
- LLaVA-PruMerge: Adaptive Token Reduction for Efficient Large Multimodal Models☆176Mar 8, 2026Updated 5 months ago