π Fastest Anything-to-Audio Gen for conditioned sound and music creation.
β245Jun 22, 2026Updated 2 months ago
Alternatives and similar repositories for AudioX-Turbo
Users that are interested in AudioX-Turbo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [SIGGRAPH 2026] Repository of Audio-Omniβ405Jun 10, 2026Updated 3 months ago
- [INTERSPEECH 2026] Pre-training, SFT, DPO and GRPO for Text-to-Audio Generationβ50Apr 17, 2026Updated 4 months ago
- β15May 8, 2021Updated 5 years ago
- β29Apr 6, 2026Updated 5 months ago
- [ICLR 2026] Repository of AudioXβ1,552Mar 10, 2026Updated 6 months ago
- Open source password manager - Proton Pass β’ AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- [ACL 2025] OZSpeech: One-step Zero-shot Speech Synthesis with Learned-Prior-Conditioned Flow Matchingβ45Feb 9, 2025Updated last year
- Foley-Omni: a unified multimodal audio generation model for task-level synthesis and complete video soundtrack generation, producing speeβ¦β27Jun 5, 2026Updated 3 months ago
- EMNLP 2026 | Official Repository of UltraVoiceβ67Aug 30, 2026Updated 2 weeks ago
- β18Jan 20, 2025Updated last year
- [EMNLP 2026' Main] HoliTok:A Continuous Holistic Tokenization with Robust Dual Capabilities of Speech Generation and Understandingβ50Jun 8, 2026Updated 3 months ago
- Ming-omni-tts: Simple and Efficient Unified Generation of Speech, Music, and Sound with Precise Controlβ265Feb 26, 2026Updated 6 months ago
- A curated list of models, benchmarks, tools and guides for audio editingβ46Updated this week
- [ACL 2026 Main] MeanAudio: Fast and Faithful Text-to-Audio Generation with Mean Flowsβ150Sep 2, 2025Updated last year
- β76Jul 29, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- β373Jul 9, 2026Updated 2 months ago
- LLaSE: Maximizing Acoustic Preservation for LLaMA based Speech Enhancementβ16Jul 11, 2025Updated last year
- β21Apr 9, 2026Updated 5 months ago
- β745Sep 5, 2026Updated last week
- β91Dec 31, 2025Updated 8 months ago
- [AAAI 2026 oral] KALL-E:Autoregressive Speech Synthesis with Next-Distribution Predictionβ43Sep 25, 2025Updated 11 months ago
- The demo page for ALMTokenizerβ59Apr 14, 2025Updated last year
- This repository contains a series of works on diffusion-based speech tokenizers, including the official implementation of the paper: "TaDβ¦β79Jan 25, 2026Updated 7 months ago
- Llasa Speed Upβ65Jan 18, 2026Updated 7 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- β50Apr 26, 2026Updated 4 months ago
- Official code for "WavCube: Unifying Speech Representation for Understanding and Generation via Semantic-Acoustic Joint Modeling"β64Jun 27, 2026Updated 2 months ago
- end-to-end text to audio scene generation modelβ51Jun 16, 2026Updated 2 months ago
- A curated list of full-duplex spoken dialogue models & benchmarksβ232Updated this week
- This is the official train-dev-test release of the Interspeech2024 Discrete Speech Representation Challenge.β32Jan 26, 2024Updated 2 years ago
- β20Oct 16, 2025Updated 10 months ago
- β61Oct 19, 2025Updated 10 months ago
- β17Apr 30, 2026Updated 4 months ago
- Unofficial PyTorch implementation of "Autoregressive Speech Synthesis without Vector Quantization (MELLE)"β41Jun 28, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available β’ AdRun AI, ML, and HPC workloads on powerful cloud GPUsβwithout limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Training-Efficient Text-to-Music Generation with State-Space Modelingβ18Jan 31, 2026Updated 7 months ago
- WavReward: Spoken Dialogue Models With Generalist Reward Evaluatorsβ57May 15, 2025Updated last year
- [NeurIPS 2025] Benchmark data and code for MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mixβ221Feb 25, 2026Updated 6 months ago
- β151Apr 25, 2025Updated last year
- This repo is text to speech with learnable audio encoder without alignment with transcript referenceβ56Sep 20, 2025Updated 11 months ago
- A trainer for SNAC (Multi-Scale Neural Audio Codec) has replaced the decoder with Vocos.β70Oct 28, 2024Updated last year
- β27Aug 29, 2025Updated last year