[NeurIPS2024] Fast and Memory-Efficient Video Diffusion Using Streamlined Inference
☆18Dec 3, 2024Updated last year
Alternatives and similar repositories for FMEDiffusion
Users that are interested in FMEDiffusion are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repository presents an efficient acceleration pipeline for Diffusion Transformer (DiT) based video generation models, optimized for …☆17Jul 13, 2026Updated last month
- Pytorch(0.4.1/1.0 verified) codes and pre-trained models for the paper: Seesaw-Net: Convolution Neural Network With Uneven Group Convolut…☆10Dec 15, 2019Updated 6 years ago
- The official Pytorch code for paper "ContextFlow: Training-Free Video Object Editing via Adaptive Context Enrichment"☆25Apr 8, 2026Updated 4 months ago
- Official implementation of Next Block Prediction: Video Generation via Semi-Autoregressive Modeling☆42Feb 12, 2025Updated last year
- 论文semantic-human-matting的代码复现☆10Jan 13, 2021Updated 5 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- [NeurIPS 2024]Taming Diffusion Prior for Image Super-Resolution with Domain Shift SDEs☆30Jun 4, 2026Updated 2 months ago
- Fast neural style with MobileNetV2 bottleneck blocks☆22May 15, 2018Updated 8 years ago
- Code for the paper "Interpreting and Improving Diffusion Models from an Optimization Perspective", appearing in ICML 2024☆15Sep 30, 2024Updated last year
- [ICCV 2025] FiVE-Bench: A Fine-grained Video Editing Benchmark for Evaluating Emerging Diffusion and Rectified Flow Models☆39Apr 2, 2026Updated 4 months ago
- The official SpeakerVid-5M data curation code.☆86Jul 23, 2025Updated last year
- Official PyTorch implementation of the paper "FlowDirector: Training-Free Flow Steering for Precise Text-to-Video Editing"☆88Dec 12, 2025Updated 8 months ago
- The implementation of MDNet, which is in submission to Interspeech2022☆14May 1, 2022Updated 4 years ago
- [AAAI 24] Official Codebase for BridgeQA: Bridging the Gap between 2D and 3D Visual Question Answering: A Fusion Approach for 3D VQA☆29Jul 12, 2024Updated 2 years ago
- Molecular Crystal Simulation Library (mcse) is an open-source Python package for manipulating and analyzing molecular crystal structures☆15May 30, 2022Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Official implementation of the paper "TOKENTRIM: INFERENCE-TIME TOKEN PRUNING FOR AUTOREGRESSIVE LONG VIDEO GENERATION"☆15Feb 8, 2026Updated 6 months ago
- [NeurIPS 2023] Sparse Modular Activation for Efficient Sequence Modeling☆40Dec 2, 2023Updated 2 years ago
- NegVSR: Augmenting Negatives for Generalized Noise Modeling in Real-world Video Super-Resolution. Real-world, video super-resolution, ima…☆12Apr 5, 2024Updated 2 years ago
- PyTorch Implementation of [AudioLCM]: a efficient and high-quality text-to-audio generation with latent consistency model.☆13Jun 15, 2024Updated 2 years ago
- ☆12Jun 9, 2025Updated last year
- [ECCV'24 Oral] PiTe: Pixel-Temporal Alignment for Large Video-Language Model☆17Feb 13, 2025Updated last year
- [NeurIPS 2024] Learning-to-Cache: Accelerating Diffusion Transformer via Layer Caching☆125Jul 15, 2024Updated 2 years ago
- ☆19Jul 14, 2023Updated 3 years ago
- ☆12Apr 18, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- a Video Quality Analysis Toolkit☆14May 16, 2025Updated last year
- Code for "A diffusion-inspired training strategy for singing voice extraction in the waveform domain" (ISMIR 2022)☆17Feb 16, 2023Updated 3 years ago
- FlowFeat: Pixel-Dense Embedding of Motion Profiles (NeurIPS 2025 Spotlight)☆117May 13, 2026Updated 3 months ago
- PGGAN implementation using TensorFlow☆10Mar 31, 2019Updated 7 years ago
- Implementation of the paper "Exploiting Time-Frequency Conformers for Music Audio Enhancement"☆16Mar 21, 2025Updated last year
- Official code release for "TDFNet: An Efficient Audio-Visual Speech Separation Model with Top-down Fusion", accepted ICIST 2023☆15Mar 17, 2024Updated 2 years ago
- ☆16Aug 10, 2025Updated last year
- ☆11Jun 2, 2019Updated 7 years ago
- [CVPR 2025 Oral] PyTorch re-implementation for Autoregressive Distillation of Diffusion Transformers (ARD).☆144Oct 1, 2025Updated 10 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆14Dec 8, 2025Updated 8 months ago
- An interactive, physically-based cloth demo in C++ using OpenGL, GLUT, GLEW, and CMake☆11Apr 15, 2022Updated 4 years ago
- This is a simple automated license plate detector developed in C++ via OpenCV.☆11Sep 26, 2020Updated 5 years ago
- This is the implementation of the manuscript "Learning General All-Neural Speech Enhancement based on Taylor's Approximation Theory", whi…☆14Nov 25, 2022Updated 3 years ago
- ☆13Dec 2, 2018Updated 7 years ago
- [ICML 2024] Compositional Image Decomposition with Diffusion Models☆55Jul 7, 2024Updated 2 years ago
- ☆20Mar 3, 2025Updated last year