Create an LJSpeech structured voice dataset on wave input
☆39Sep 28, 2024Updated last year
Alternatives and similar repositories for Audio-to-Voice-Dataset
Users that are interested in Audio-to-Voice-Dataset are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- German small and large versions of GPT2.☆20May 11, 2022Updated 4 years ago
- Loads OpenSubtitles v2018 dataset without having to load everything into memory at once. Works well with pytorch.☆13Aug 26, 2020Updated 6 years ago
- notes on langchain☆18Mar 20, 2026Updated 5 months ago
- TTS Client for Coqui TTS server☆13Jan 7, 2023Updated 3 years ago
- A Simple Flask App to interact with your Machine Translation Model☆13Feb 26, 2020Updated 6 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [EMNLP Main '25] LiteASR: Efficient Automatic Speech Recognition with Low-Rank Approximation☆157May 18, 2025Updated last year
- A Diffrentiable WFST-based End-to-End Automatic Speech Recognition toollkit with flexible topology support☆12Feb 15, 2026Updated 6 months ago
- Toy example to illustrate how to use kaldi recipes.☆13Mar 11, 2021Updated 5 years ago
- Jax, Flax, examples (ImageClassification, SemanticSegmentation, and more...)☆10May 10, 2025Updated last year
- An ASR toolkit with the freedom of topology☆10Dec 18, 2023Updated 2 years ago
- Clean interface for localStorage / sessionStorage☆15Apr 13, 2017Updated 9 years ago
- Automatic Speech Recognition (ASR) system for the Samrómur speech corpus using Kaldi☆12Sep 30, 2022Updated 3 years ago
- Adapted OS for e-ink tablets - allows to use work-related apps with no harm for eyes☆11May 17, 2020Updated 6 years ago
- An educational tool to train, inspect, evaluate and translate using neural engines☆21Mar 13, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- VoiceRestore: Flow-Matching Transformers for Universal Speech Restoration☆209Apr 21, 2025Updated last year
- Run Retrieval-based Voice Conversion training and inference with ease.☆12Jan 24, 2025Updated last year
- A highly optimized engine for neutts-air model to generate minutes of audio in seconds. Over 200x realtime on modern hardware!☆120Nov 24, 2025Updated 9 months ago
- Search for PII in Python☆29Jan 29, 2024Updated 2 years ago
- DefCon Red Team Village 2023 Workshop on DLL Sideloading☆19Aug 15, 2023Updated 3 years ago
- More than Just Words: Modeling Non-textual Characteristics of Podcasts☆26Nov 6, 2019Updated 6 years ago
- A collection of utilities for handling IPA phones.☆27Sep 24, 2023Updated 2 years ago
- Repository containing scripts/helpers for configuring a Raspberry Pi to work with XMOS mic frontend☆14Jul 31, 2023Updated 3 years ago
- Safari Reader Mode Source Code☆22Mar 5, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆81Mar 12, 2026Updated 5 months ago
- Enable RNNLM lattice rescoring with Pytorch [kaldi]☆12Jun 5, 2020Updated 6 years ago
- High accuracy code-switching whisper / qwen3 transcription☆42Jun 17, 2026Updated 2 months ago
- ☆401Sep 3, 2024Updated last year
- ☆19Sep 9, 2024Updated last year
- Towards Efficient and Multifaceted Computer-assisted Pronunciation Training Leveraging Hierarchical Selective State Space Model and Decou…☆16May 6, 2025Updated last year
- Deep Speech Distances PyTorch☆29Feb 21, 2022Updated 4 years ago
- Second SIGMORPHON Shared Task on Grapheme-to-Phoneme Conversions☆25Jun 7, 2021Updated 5 years ago
- Thorsten-Voice: A free to use, offline working, high quality german TTS voice should be available for every project without any license s…☆729Aug 4, 2026Updated 3 weeks ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A BiRNN framework implemented in Python and TensorFlow to extract parallel sentences from aligned comparable corpora.☆33Sep 4, 2018Updated 7 years ago
- llmware RAG Demo App.☆17Dec 10, 2023Updated 2 years ago
- High-performance Google Colab Notebook for fast & accurate audio transcription/translation using OpenAI Whisper. Accelerated on TPUs with…☆18Jun 8, 2025Updated last year
- Reactive Multi-language Gradio App with minimal effort☆21Oct 12, 2025Updated 10 months ago
- ☆12Jan 2, 2024Updated 2 years ago
- Dia-JAX: A JAX port of Dia, the text-to-speech model for generating realistic dialogue from text with emotion and tone control.☆30May 7, 2025Updated last year
- ☆32Jun 30, 2023Updated 3 years ago