☆33Nov 27, 2021Updated 4 years ago
Alternatives and similar repositories for wenet_stt_python
Users that are interested in wenet_stt_python are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Wenet speech to text for react native☆10Nov 1, 2022Updated 3 years ago
- Went online decode demo☆31Apr 28, 2021Updated 5 years ago
- ☆40Aug 15, 2021Updated 4 years ago
- ☆50Dec 26, 2020Updated 5 years ago
- This will hold the crowdsourcing platform to be used to store voice data from various speakers which will act as input dataset for speech…☆17Mar 6, 2023Updated 3 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Implementation of different noise embeddings for noise aware training of Kaldi acoustic models.☆13Feb 13, 2021Updated 5 years ago
- ☆61Jan 31, 2023Updated 3 years ago
- ☆18Apr 28, 2021Updated 5 years ago
- Properly handle position-dependent phones in a subword lexicon FST☆31Oct 26, 2020Updated 5 years ago
- Speechflow for emotion recognition related information decomposition☆10Jul 27, 2021Updated 5 years ago
- This app is intended to automatically create a corpus for ASR systems using pseudo-labeling.☆27Feb 15, 2024Updated 2 years ago
- A free & open tool for transcribing audio interviews with offline ASR support☆25Dec 21, 2023Updated 2 years ago
- Finally, some decent sample sentences☆24Dec 3, 2023Updated 2 years ago
- Towards hot directions in industrial end to end speech recognition☆329Nov 30, 2021Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆33Aug 6, 2021Updated 5 years ago
- ☆17Apr 14, 2023Updated 3 years ago
- CAT is more than a CRF-based ASR toolkit: it provides a complete workflow for data-efficient end-to-end ASR, supporting CTC, CTC-CRF, RNN…☆368Feb 5, 2026Updated 6 months ago
- ☆43Jun 25, 2018Updated 8 years ago
- This is a mirror of https://gitlab.com/tiro-is/tiro-speech-core☆15Jun 19, 2023Updated 3 years ago
- Docker image and scripts for training finetuned or completely personal Kaldi speech models. Particularly for use with kaldi-active-gramma…☆21Jan 24, 2022Updated 4 years ago
- List of NN based singal processing papers☆23Jun 5, 2023Updated 3 years ago
- A repository for Chinese text normalization.☆20May 2, 2021Updated 5 years ago
- Kaldi model converter to ONNX☆248Jan 27, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Discriminative Neural Clustering for Speaker Diarisation☆79Apr 8, 2022Updated 4 years ago
- ☆13Mar 25, 2021Updated 5 years ago
- 3M: Multi-loss, Multi-path and Multi-level Neural Networks for speech recognition☆119Jun 22, 2022Updated 4 years ago
- An online speech recognition extension toolkit of Kaldi☆55Jun 23, 2021Updated 5 years ago
- Torch-based tool for quantizing high-dimensional vectors using additive codebooks☆54May 25, 2022Updated 4 years ago
- Keyword Search Recipe for Subword ASR☆30Jul 12, 2019Updated 7 years ago
- Pronunciation lexicon covering both English and Chinese languages for Automatic Speech Recognition.☆263Oct 11, 2019Updated 6 years ago
- finite-state toolkit, EM and Bayesian (Gibbs sampling) training for FST and context-free derivation forests☆15Jan 24, 2017Updated 9 years ago
- Online streaming speaker change detection model in Pytorch☆44Apr 14, 2023Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- The People’s Speech Dataset☆115Jan 11, 2024Updated 2 years ago
- Deepspeech ASR Model for the Catalan Language☆17Feb 15, 2021Updated 5 years ago
- My solution to course E6870 (Speech Recognition) of Columbia University.☆37May 13, 2018Updated 8 years ago
- Implementation of the DIVA model of speech acquisition and production using PyTorch☆23Jan 18, 2023Updated 3 years ago
- ☆30Jul 21, 2022Updated 4 years ago
- Efficient Neural Architecture Search via Straight-Through Gradients☆13Nov 12, 2020Updated 5 years ago
- A KALDI/C++ implementation of GoogleBrain's SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition☆15Sep 4, 2019Updated 6 years ago