Annotated Flow Matching paper
☆235Sep 14, 2024Updated 2 years ago
Alternatives and similar repositories for flow-matching
Users that are interested in flow-matching are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Educational implementation of the Discrete Flow Matching paper☆141Aug 26, 2024Updated 2 years ago
- TorchCFM: a Conditional Flow Matching library☆2,585Jul 20, 2026Updated 2 months ago
- A summary of related works about flow matching, stochastic interpolants☆697Apr 12, 2026Updated 5 months ago
- A PyTorch library for implementing flow matching algorithms, featuring continuous and discrete flow matching implementations. It includes…☆4,758Jan 5, 2026Updated 8 months ago
- [INTERSPEECH 2025 Oral]Official code for "Accelerating Diffusion-based Text-to-Speech Model Training with Dual Modality Alignment"☆68Jun 16, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Official PyTorch implementation of the paper: Flow Matching in Latent Space☆364Jan 20, 2025Updated last year
- ☆14May 21, 2024Updated 2 years ago
- Adaptation of vision-language models (CLIP) to downstream tasks using local and global prompts.☆53Jul 10, 2025Updated last year
- [WAVC 2024] Official implementation of the paper: Semantic Generative Augmentations for Few-shot Counting☆13May 1, 2024Updated 2 years ago
- Official Implementation of Rectified Flow (ICLR2023 Spotlight)☆1,648Jul 20, 2024Updated 2 years ago
- Implementation of rectified flow and some of its followup research / improvements in Pytorch☆486Sep 12, 2026Updated last week
- ☆27Feb 8, 2025Updated last year
- LAFMA: A Latent Flow Matching Model for Text-to-Audio Generation (INTERSPEECH 2024)☆44Jun 13, 2024Updated 2 years ago
- Unofficial implementation of Variational Diffusion Models in PyTorch (Lightning)☆12Aug 31, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for CTC-TTS: LLM-based dual-streaming text-to-speech with CTC alignment, Interspeech 2026.☆22Aug 14, 2026Updated last month
- Normalizing flows in PyTorch☆466Mar 10, 2026Updated 6 months ago
- Meanflow and multilingual for F5-TTS model☆16Aug 23, 2025Updated last year
- [ACL 2026 Main] MeanAudio: Fast and Faithful Text-to-Audio Generation with Mean Flows☆150Sep 2, 2025Updated last year
- Aligned Diffusion Schroedinger Bridges (UAI 2023)☆14Sep 18, 2025Updated last year
- [ICCV'25] ViLU: Learning Vision-Language Uncertainties for Failure Prediction☆16Jul 16, 2025Updated last year
- Implementation of Action Matching for the Schrödinger equation☆25Jun 18, 2023Updated 3 years ago
- Implementation of a single layer of the MMDiT, proposed in Stable Diffusion 3, in Pytorch☆561Jan 18, 2026Updated 8 months ago
- Official code for "F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization"☆169Mar 3, 2026Updated 6 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Official Implementation for "Consistency Flow Matching: Defining Straight Flows with Velocity Consistency"☆272Jan 17, 2025Updated last year
- ICML 2023: Reduce, Reuse, Recycle: Composing Energy-Based Diffusion Models with MCMC☆151Oct 18, 2024Updated last year
- DiffPhase: Generative Diffusion-based STFT Phase Retrieval☆16Sep 21, 2023Updated 3 years ago
- Code for ICLR 2024 Paper: CompA: Addressing the Gap in Compositional Reasoning in Audio-Language Models☆23Jul 10, 2024Updated 2 years ago
- Speech Resynthesis and Language Modeling☆27Jun 11, 2025Updated last year
- Make-An-Audio-3: Transforming Text/Video into Audio via Flow-based Large Diffusion Transformers☆121May 19, 2025Updated last year
- Phoneme-Level BERT for Enhanced Prosody of Text-to-Speech with Grapheme Predictions☆269Jan 13, 2025Updated last year
- ☆82Aug 11, 2025Updated last year
- Inference codebase for "Cacophony: An Improved Contrastive Audio-Text Model". Preprint: https://arxiv.org/abs/2402.06986☆49Jan 19, 2026Updated 8 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆12May 23, 2018Updated 8 years ago
- ☆13Dec 12, 2023Updated 2 years ago
- Implementation of papers in 101 lines of code.☆18Nov 12, 2023Updated 2 years ago
- PyTorch implementation of MeanFlow & iMF (one-step generative modeling).☆1,202Jul 1, 2026Updated 2 months ago
- An official pytorch implementation of EACL2024 short paper "Flow Matching for Conditional Text Generation in a Few Sampling Steps"☆34Jul 17, 2025Updated last year
- ☆57Jul 16, 2025Updated last year
- Official PyTorch Implementation of "SiT: Exploring Flow and Diffusion-based Generative Models with Scalable Interpolant Transformers"☆1,214Dec 22, 2025Updated 9 months ago