Repository of lessons exploring image diffusion models, focused on understanding and education.
☆64Jan 9, 2025Updated last year
Alternatives and similar repositories for mindiffusion
Users that are interested in mindiffusion are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repo implements Video generation model using Latent Diffusion Transformers(Latte) in PyTorch and provides training and inference cod…☆19Jan 6, 2025Updated last year
- ☆20Sep 11, 2024Updated last year
- ☆17Jan 22, 2025Updated last year
- Implemented a stable diffusion architecture using PyTorch.☆91Jan 3, 2024Updated 2 years ago
- ☆16Jun 14, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Create bounding boxes selecting masks by threshold.☆31May 22, 2024Updated 2 years ago
- ☆16Mar 1, 2022Updated 4 years ago
- ☆14Oct 14, 2024Updated last year
- FlexiFilm: Long Video Generation with Flexible Conditions☆31May 1, 2024Updated 2 years ago
- This repository is for The Power of Sound(TPoS): Audio Reactive Video Generation with Stable Diffusion (ICCV2023)☆25Dec 7, 2023Updated 2 years ago
- implementation of https://arxiv.org/pdf/2312.09299☆21Jul 3, 2024Updated 2 years ago
- ☆12Dec 14, 2024Updated last year
- A one-stop library to standardize the inference and evaluation of all the conditional video generation models.☆50Feb 13, 2025Updated last year
- OpenAI Whisper demo on Axera☆17Jan 15, 2026Updated 6 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [BMVC 2023] Zero-shot Composed Text-Image Retrieval☆55Nov 26, 2024Updated last year
- ☆14Oct 31, 2023Updated 2 years ago
- A unified media (Image, Video, Audio, Text) diffusion repository, for education and learning.☆47Apr 5, 2025Updated last year
- ☆15May 26, 2026Updated 2 months ago
- EmoCAST: Emotional Talking Portrait via Emotive Text Description☆35Dec 23, 2025Updated 7 months ago
- This is a series of notebooks to support lectures on Time series analysis and forecast for a course I held in a master postgraduate progr…☆15Nov 29, 2022Updated 3 years ago
- Kaldi code for doing DNN with tensorflow☆13Feb 8, 2016Updated 10 years ago
- Implementation of E2-TTS, "Embarrassingly Easy Fully Non-Autoregressive Zero-Shot TTS", in MLX☆21Oct 8, 2024Updated last year
- ☆23Jan 23, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Conformer block with Rotary Position Embedding, modified from lucidrains' implement☆19Sep 13, 2024Updated last year
- Fastest Stable Diffusion GUI, Pipeline, with only one python script, with the least number of lines and in the least complex way.☆13Jan 8, 2025Updated last year
- Experiment with NVIDIA Triton and Whisper☆15Apr 29, 2024Updated 2 years ago
- ☆190Dec 23, 2024Updated last year
- Official code for ICCV 2023 paper: "Efficient Emotional Adaptation for Audio-Driven Talking-Head Generation".☆300Mar 4, 2026Updated 5 months ago
- C++ implementation of the enthalpy-based thermal evolution of loops (EBTEL) model wrapped in Python☆13Apr 10, 2026Updated 4 months ago
- Source code of the TextLap model, a LLM for text-2-layout generation.☆18Oct 21, 2024Updated last year
- ☆15May 13, 2024Updated 2 years ago
- My learnings (publicly) on RAG systems☆15Jan 2, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Train the next generation of TTS systems.☆169Sep 13, 2024Updated last year
- Simple model creation for time series classification in Pytorch☆13Aug 24, 2025Updated 11 months ago
- Somewhat janky implementation of Sonar Sampling (momentum based sampling) for ComfyUI along with an assortment of advanced noise tools (s…☆49Aug 7, 2026Updated last week
- user interface for AI image generation☆15Dec 26, 2022Updated 3 years ago
- Official implementation of "Divide & Bind Your Attention for Improved Generative Semantic Nursing" (BMVC 2023 Oral)☆38Jan 25, 2024Updated 2 years ago
- ChatGPT-based voice Telegram bot☆17Dec 17, 2024Updated last year
- [BMVC'24] G3FA: Geometry-guided GAN for Face Animation☆20Mar 14, 2025Updated last year