[NeurIPS2024] Fast and Memory-Efficient Video Diffusion Using Streamlined Inference
☆17Dec 3, 2024Updated last year
Alternatives and similar repositories for FMEDiffusion
Users that are interested in FMEDiffusion are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repository presents an efficient acceleration pipeline for Diffusion Transformer (DiT) based video generation models, optimized for …☆17Jul 13, 2026Updated 2 months ago
- The official Pytorch code for paper "ContextFlow: Training-Free Video Object Editing via Adaptive Context Enrichment"☆24Apr 8, 2026Updated 5 months ago
- [ICLR 2026] Official Implementation of “LightCtrl: Training-free Controllable Video Relighting”☆30Mar 31, 2026Updated 6 months ago
- [NeurIPS 2024]Taming Diffusion Prior for Image Super-Resolution with Domain Shift SDEs☆30Jun 4, 2026Updated 3 months ago
- Code for the paper "Interpreting and Improving Diffusion Models from an Optimization Perspective", appearing in ICML 2024☆15Sep 30, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [NeurIPS 2022] Latency-aware Spatial-wise Dynamic Networks☆25Aug 21, 2023Updated 3 years ago
- ☆14Mar 8, 2026Updated 6 months ago
- The official SpeakerVid-5M data curation code.☆88Jul 23, 2025Updated last year
- 语音合成服务☆12Mar 18, 2023Updated 3 years ago
- [ICCV 2025] FiVE-Bench: A Fine-grained Video Editing Benchmark for Evaluating Emerging Diffusion and Rectified Flow Models☆39Apr 2, 2026Updated 6 months ago
- The implementation of MDNet, which is in submission to Interspeech2022☆14May 1, 2022Updated 4 years ago
- Official PyTorch implementation of the paper "FlowDirector: Training-Free Flow Steering for Precise Text-to-Video Editing"☆88Dec 12, 2025Updated 9 months ago
- Official implementation of the paper "TOKENTRIM: INFERENCE-TIME TOKEN PRUNING FOR AUTOREGRESSIVE LONG VIDEO GENERATION"☆15Feb 8, 2026Updated 7 months ago
- PyTorch Implementation of [AudioLCM]: a efficient and high-quality text-to-audio generation with latent consistency model.☆13Jun 15, 2024Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆12Feb 8, 2024Updated 2 years ago
- ☆12Jun 9, 2025Updated last year
- [ECCV'24 Oral] PiTe: Pixel-Temporal Alignment for Large Video-Language Model☆17Feb 13, 2025Updated last year
- [NeurIPS 2024] Learning-to-Cache: Accelerating Diffusion Transformer via Layer Caching☆125Jul 15, 2024Updated 2 years ago
- ☆12Apr 18, 2025Updated last year
- a Video Quality Analysis Toolkit☆14May 16, 2025Updated last year
- PGGAN implementation using TensorFlow☆10Mar 31, 2019Updated 7 years ago
- Implementation of the paper "Exploiting Time-Frequency Conformers for Music Audio Enhancement"☆16Mar 21, 2025Updated last year
- Official code release for "TDFNet: An Efficient Audio-Visual Speech Separation Model with Top-down Fusion", accepted ICIST 2023☆15Mar 17, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆11Jun 2, 2019Updated 7 years ago
- [CVPR 2025 Oral] PyTorch re-implementation for Autoregressive Distillation of Diffusion Transformers (ARD).☆144Oct 1, 2025Updated last year
- ☆14Dec 8, 2025Updated 9 months ago
- An interactive, physically-based cloth demo in C++ using OpenGL, GLUT, GLEW, and CMake☆11Apr 15, 2022Updated 4 years ago
- This is the implementation of the manuscript "Learning General All-Neural Speech Enhancement based on Taylor's Approximation Theory", whi…☆14Nov 25, 2022Updated 3 years ago
- ☆13Dec 2, 2018Updated 7 years ago
- [ICML 2024] Compositional Image Decomposition with Diffusion Models☆55Jul 7, 2024Updated 2 years ago
- ☆20Mar 3, 2025Updated last year
- Make any person bald!! Component of the paper: Learning to regulate 3D head shape by removing occluding hair from in-the-wild images.☆12Jun 6, 2022Updated 4 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Speech Separation☆11Jan 6, 2022Updated 4 years ago
- AMD 0.9B efficient text to video diffusion model☆48May 16, 2026Updated 4 months ago
- Scalable group inference for generating high quality and diverse images with diffusion models.☆43Aug 31, 2025Updated last year
- Train a Jev-like multimodal model by yourself. System One Model, now with vision.☆561Updated this week
- ☆20Jun 22, 2026Updated 3 months ago
- ☆14Jan 12, 2023Updated 3 years ago
- Official Pytorch implementation for "AttentionHand: Text-driven Controllable Hand Image Generation for 3D Hand Reconstruction in the Wild…☆12May 11, 2026Updated 4 months ago