This code implements a basic MLP for speech recognition. The MLP is trained with pytorch, while feature extraction, alignments, and decoding are performed with Kaldi. The current implementation supports dropout and batch normalization. An example for phoneme recognition using the standard TIMIT dataset is provided.
☆40Feb 10, 2018Updated 8 years ago
Alternatives and similar repositories for pytorch_MLP_for_ASR
Users that are interested in pytorch_MLP_for_ASR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Empirical Evaluation of Speaker Adaptation on DNN based Acoustic Model☆13Nov 25, 2019Updated 6 years ago
- ☆16May 25, 2019Updated 7 years ago
- ☆24Sep 25, 2018Updated 7 years ago
- Baseline kaldi script for UA-SPEECH corpus☆32Oct 16, 2024Updated last year
- Python implementation of pre-processing for End-to-End speech recognition☆70Feb 19, 2018Updated 8 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Extended speech recognition neural network based on Kaldi for reproducible research☆15Aug 28, 2015Updated 11 years ago
- Phone-level evaluation of L2 speakers (GOP algorithm)☆27Mar 1, 2017Updated 9 years ago
- PyTorch bindings for Warp-CTC☆42Dec 6, 2019Updated 6 years ago
- Custom decoders for Kaldi☆13Jun 5, 2019Updated 7 years ago
- ☆13Sep 12, 2017Updated 9 years ago
- implement end-to-end asr algorithm with tensorflow☆40Aug 23, 2018Updated 8 years ago
- speech-aligner,是一个从“人声语音”及其“语言文本”,产生音素级别时间对齐标注的工具。speech-aligner, is a tool that generate phoneme-level alignment between human speech an…☆15Dec 19, 2018Updated 7 years ago
- it's ASR decoder and make graph project☆33May 26, 2022Updated 4 years ago
- Tensor2tensor experiment with SpecAugment☆46May 13, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Coordinate-wise meta-learner for speaker adaptation of ASR models.☆20Dec 30, 2019Updated 6 years ago
- Conda based Python distribution for the Raspberry Pi☆11Jul 16, 2018Updated 8 years ago
- A library of speech gadgets.☆16Oct 15, 2022Updated 3 years ago
- ☆43Jun 25, 2018Updated 8 years ago
- Extract mfcc vectors and phones from TIMIT dataset☆17Mar 23, 2023Updated 3 years ago
- Stochastic Downsampling for Cost-Adjustable Inference and Improved Regularization in Convolutional Networks☆18Nov 5, 2019Updated 6 years ago
- solutions for https://www.kaggle.com/c/tensorflow-speech-recognition-challenge☆31Jan 28, 2018Updated 8 years ago
- Semi-supervised spoken language understanding (SLU) via self-supervised speech and language model pretraining☆12Mar 23, 2021Updated 5 years ago
- A SPMI Lab toolkit for language models.☆11Apr 12, 2017Updated 9 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- 语音识别 语音前端处理 语音合成 语音转换等等语音技术的资料汇总☆23Nov 8, 2019Updated 6 years ago
- It shows how to build and optimize Decision Tree Classifier of "Diabetes dataset" using Python Scikit-learn package.☆16Aug 21, 2020Updated 6 years ago
- pytorch-kaldi is a project for developing state-of-the-art DNN/RNN hybrid speech recognition systems. The DNN part is managed by pytorch,…☆2,406Mar 14, 2022Updated 4 years ago
- Keyword Spotting suitable for embedded devices.☆28Jun 22, 2020Updated 6 years ago
- SWIG bindings for Kaldi I/O, built with Conda☆15Dec 15, 2024Updated last year
- FastLongSpeech is a novel framework designed to extend the capabilities of Large Speech-Language Models for efficient long-speech process…☆16Jul 22, 2025Updated last year
- Python-based cross-platform tool for mining text data (html, transcript, problems) of edX MOOCs on a user's dashboard. It is an extension…☆10Feb 12, 2020Updated 6 years ago
- python wrap for hts engine☆14Jan 30, 2018Updated 8 years ago
- This repository can be used to perform Speech to Text Conversion in multiple Languages, e.g., It can convert whatever you are speaking in…☆11Oct 6, 2020Updated 5 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Listen, Attend and Spell - PyTorch Implementation☆17Dec 28, 2018Updated 7 years ago
- ☆13Feb 8, 2017Updated 9 years ago
- A library that you can use to build spiking neural network brains for your Arduino robots! Largely allows you to follow the paradigms of …☆10Nov 21, 2015Updated 10 years ago
- Bidirectional dynamic RNN + CTC for phoneme recognition☆47Jun 24, 2020Updated 6 years ago
- PyTorch Implementations for End-to-End Automatic Speech Recognition☆127Jun 10, 2019Updated 7 years ago
- Telluride neuromorphic project on event-based image classification☆11Dec 21, 2016Updated 9 years ago
- An emulator of General Vision's CM1K neuromorphic chip, including modeling experiments using several popular public datasets☆10Jun 2, 2017Updated 9 years ago