this is a treasure-house of speech
☆168Jun 25, 2018Updated 8 years ago
Alternatives and similar repositories for awesome-speech
Users that are interested in awesome-speech are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A Python interface to OpenFst (fix FstDrawer interface issue for 1.6 version)☆17Apr 2, 2018Updated 8 years ago
- This is an implementation of "Generative adversarial network-based postfilter for statistical parametric speech synthesis"☆16Jun 27, 2018Updated 8 years ago
- Custom decoders for Kaldi☆81Jun 10, 2019Updated 7 years ago
- End-to-end ASR/LM implementation with PyTorch☆595Aug 30, 2021Updated 4 years ago
- DaCiDian is an open-sourced chinese mandarin lexicon for automatic speech recognition(ASR)☆301Jun 15, 2020Updated 6 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Working online speech recognition based on RNN Transducer. ( Trained model release available in release )☆292Aug 5, 2021Updated 5 years ago
- Custom decoders for Kaldi☆13Jun 5, 2019Updated 7 years ago
- Pronunciation lexicon covering both English and Chinese languages for Automatic Speech Recognition.☆263Oct 11, 2019Updated 6 years ago
- This is a list of features, scripts, blogs and resources for better using Kaldi ( http://kaldi-asr.org/ )☆536Feb 9, 2022Updated 4 years ago
- Interspeech 2019 tutorial materials☆49Sep 26, 2019Updated 6 years ago
- ASR with PyTorch☆139Mar 10, 2019Updated 7 years ago
- ☆76Mar 18, 2022Updated 4 years ago
- An open-source speech separation and enhancement library☆214May 13, 2020Updated 6 years ago
- INTERSPEECH 2019 Tutorial Materials☆194Mar 30, 2021Updated 5 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Contains code for our work on speech to singing conversion (ICASSP 2020)☆50Oct 27, 2020Updated 5 years ago
- This is an extension of kaldi speech recognition software which allows to perform decoding of speech with hybrid word and phoneme graphs.…☆11Feb 4, 2020Updated 6 years ago
- ☆24Mar 13, 2020Updated 6 years ago
- Unsupervised speech activity detection system.☆11Jul 2, 2018Updated 8 years ago
- This repo contains the baseline model recipes and pre-trained model for GramVanni hindi ASR challenge☆16Mar 26, 2022Updated 4 years ago
- BERT and LSTM baseline models of the ZeroSpeech Challenge 2021☆60Oct 19, 2022Updated 3 years ago
- Convert words to numbers☆21Apr 13, 2022Updated 4 years ago
- it's ASR decoder and make graph project☆33May 26, 2022Updated 4 years ago
- Chinese text normalization for speech processing☆737Mar 18, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Minimize kaldi nnet3 chain decoder☆45Jan 10, 2020Updated 6 years ago
- ☆57Oct 6, 2021Updated 4 years ago
- scripts to align a given wave to its transcription using trained models by Kaldi☆37Aug 15, 2019Updated 7 years ago
- A Pytorch Implementation of Transducer Model for End-to-End Speech Recognition☆238May 12, 2020Updated 6 years ago
- A collection of examples demonstrating how we can build speech synthesis systems using nnmnkwii.☆70May 15, 2020Updated 6 years ago
- An Automatic Speech Recognition using GMM & HMM.☆19Aug 16, 2019Updated 6 years ago
- Tensorflow version of DFSMN☆49Jul 17, 2018Updated 8 years ago
- Towards hot directions in industrial end to end speech recognition☆329Nov 30, 2021Updated 4 years ago
- Crystal - C++ implementation of a unified framework for multilingual TTS synthesis engine with SSML specification as interface.☆230Aug 17, 2020Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- CUDA-Warp RNN-Transducer☆215Feb 22, 2023Updated 3 years ago
- Self-Supervised Speech Pre-training and Representation Learning Toolkit☆2,561Mar 12, 2026Updated 5 months ago
- A PyTorch implementation of Speech Transformer, an End-to-End ASR with Transformer network on Mandarin Chinese.☆810Apr 6, 2023Updated 3 years ago
- Yet another speech toolkit based on Kaldi and PyTorch☆173Jul 1, 2020Updated 6 years ago
- This repository provides data and code for "Vox Populi, Vox DIY: Benchmark Dataset for Crowdsourced Audio Transcription" paper.☆16Jul 22, 2021Updated 5 years ago
- ASR cases for speech handbook at CSLT-THU, based on Kaldi toolkit and Thchs30 database, in egs/cslt_cases.☆107Mar 12, 2021Updated 5 years ago
- ☆277Jan 15, 2021Updated 5 years ago