cumulo-autumn/StreamDiffusion

Readme badge preview -

If you own this repo, copy the snippet below and add it to your README.md

[![RelatedRepos](https://img.shields.io/badge/related-repos-yellow)](https://relatedrepos.com/gh/cumulo-autumn/StreamDiffusion)

cumulo-autumn / StreamDiffusion

StreamDiffusion: A Pipeline-Level Solution for Real-Time Interactive Generation

☆10,780

Alternatives and similar repositories for StreamDiffusion

Users that are interested in StreamDiffusion are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.

Sorting:

guoyww / AnimateDiff
View on GitHub
Official implementation of AnimateDiff.
☆12,187Jul 31, 2024Updated last year
TencentARC / PhotoMaker
View on GitHub
PhotoMaker [CVPR 2024]
☆10,097Oct 31, 2024Updated last year
ali-vilab / AnyDoor
View on GitHub
Official implementations for paper: Anydoor: zero-shot object-level image customization
☆4,230Apr 8, 2024Updated 2 years ago
tyxsspa / AnyText
View on GitHub
Official implementation code of the paper <AnyText: Multilingual Visual Text Generation And Editing>
☆4,863Mar 7, 2025Updated last year
luosiallen / latent-consistency-model
View on GitHub
Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference
☆4,614Jun 14, 2024Updated 2 years ago
GPUs on demand by Runpod - Special Offer Available • Ad
Run AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
magic-research / magic-animate
View on GitHub
[CVPR 2024] Official repository for "MagicAnimate: Temporally Consistent Human Image Animation using Diffusion Model"
☆10,901Aug 29, 2025Updated 10 months ago
Stability-AI / generative-models
View on GitHub
Generative Models by Stability AI
☆27,221Dec 16, 2025Updated 7 months ago
apple / ml-ferret
View on GitHub
☆8,674Oct 9, 2024Updated last year
instantX-research / InstantID
View on GitHub
InstantID: Zero-shot Identity-Preserving Generation in Seconds 🔥
☆11,968Jul 18, 2024Updated 2 years ago
ali-vilab / VGen
View on GitHub
Official repo for VGen: a holistic video generation ecosystem for video generation building on diffusion models
☆3,156Jan 10, 2025Updated last year
HumanAIGC / AnimateAnyone
View on GitHub
Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation
☆14,790Sep 20, 2025Updated 9 months ago
Stability-AI / StableCascade
View on GitHub
Official Code for Stable Cascade
☆6,545Jul 25, 2024Updated last year
myshell-ai / OpenVoice
View on GitHub
Instant voice cloning by MIT and MyShell. Audio foundation model.
☆36,943Apr 19, 2025Updated last year
Comfy-Org / ComfyUI
View on GitHub
The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.
☆120,828Updated this week
Managed Kubernetes at scale on DigitalOcean • Ad
DigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
hpcaitech / Open-Sora
View on GitHub
Open-Sora: Democratizing Efficient Video Production for All
☆29,186Apr 9, 2026Updated 3 months ago
guoqincode / Open-AnimateAnyone
View on GitHub
Unofficial Implementation of Animate Anyone
☆2,925Jul 9, 2024Updated 2 years ago
lllyasviel / ControlNet
View on GitHub
Let us control diffusion models!
☆34,000Feb 25, 2024Updated 2 years ago
lllyasviel / Omost
View on GitHub
Your image is almost there!
☆7,611Jul 26, 2024Updated last year
tencent-ailab / IP-Adapter
View on GitHub
The image prompt adapter is designed to enable a pretrained text-to-image diffusion model to generate images with image prompt.
☆6,630Jun 28, 2024Updated 2 years ago
PKU-YuanGroup / Open-Sora-Plan
View on GitHub
This project aim to reproduce Sora (Open AI T2V model), we wish the open source community contribute to this project.
☆12,154Mar 8, 2026Updated 4 months ago
YangLing0818 / RPG-DiffusionMaster
View on GitHub
[ICML 2024] Mastering Text-to-Image Diffusion: Recaptioning, Planning, and Generating with Multimodal LLMs (RPG)
☆1,839Feb 1, 2025Updated last year
lllyasviel / Fooocus
View on GitHub
Focus on prompting and generating
☆51,092Dec 1, 2025Updated 7 months ago
modelscope / DiffSynth-Studio
View on GitHub
Enjoy the magic of Diffusion models!
☆12,706Updated this week
Deploy on Railway without the complexity - Free Credits Offer • Ad
Connect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
AILab-CVC / VideoCrafter
View on GitHub
VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models
☆5,066Jan 9, 2026Updated 6 months ago
HVision-NKU / StoryDiffusion
View on GitHub
Accepted as [NeurIPS 2024] Spotlight Presentation Paper
☆6,441Sep 26, 2024Updated last year
MooreThreads / Moore-AnimateAnyone
View on GitHub
Character Animation (AnimateAnyone, Face Reenactment)
☆3,511May 31, 2024Updated 2 years ago
modelscope / facechain
View on GitHub
FaceChain is a deep-learning toolchain for generating your Digital-Twin.
☆9,500Jun 6, 2025Updated last year
facefusion / facefusion
View on GitHub
Industry leading face manipulation platform
☆29,288Updated this week
facebookresearch / audiocraft
View on GitHub
Audiocraft is a library for audio processing and generation with deep learning. It features the state-of-the-art EnCodec audio compressor…
☆23,476Mar 3, 2026Updated 4 months ago
suno-ai / bark
View on GitHub
🔊 Text-Prompted Generative Audio Model
☆39,197Aug 19, 2024Updated last year
facebookresearch / audio2photoreal
View on GitHub
Code and dataset for photorealistic Codec Avatars driven from audio
☆2,854Sep 15, 2024Updated last year
Tiiny-AI / PowerInfer
View on GitHub
High-speed Large Language Model Serving for Local Deployment
☆9,636May 11, 2026Updated 2 months ago
Deploy to Railway using AI coding agents - Free Credits Offer • Ad
Use Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
dreamoving / dreamoving-project
View on GitHub
Official implementation of DreaMoving
☆1,790Jan 9, 2024Updated 2 years ago
lllyasviel / IC-Light
View on GitHub
More relighting!
☆8,468Feb 20, 2025Updated last year
radames / Real-Time-Latent-Consistency-Model
View on GitHub
App showcasing multiple real-time diffusion models pipelines with Diffusers
☆915Sep 27, 2025Updated 9 months ago
haotian-liu / LLaVA
View on GitHub
[NeurIPS'23 Oral] Visual Instruction Tuning (LLaVA) built towards GPT-4V level capabilities and beyond.
☆24,923Aug 12, 2024Updated last year
open-mmlab / Amphion
View on GitHub
Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junio…
☆9,931Mar 25, 2026Updated 3 months ago
AUTOMATIC1111 / stable-diffusion-webui
View on GitHub
Stable Diffusion web UI
☆164,252Mar 2, 2026Updated 4 months ago
LargeWorldModel / LWM
View on GitHub
Large World Model -- Modeling Text and Video with Millions Context
☆7,427Oct 19, 2024Updated last year