Project to learn about speech recognition - both Speaker Diarization and other Speech Recognition applications.
☆50Feb 1, 2017Updated 9 years ago
Alternatives and similar repositories for Deep-Learning-Speech-Recognition
Users that are interested in Deep-Learning-Speech-Recognition are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Unsupervised Speaker Clustering & Speaker Recognition☆13Jan 7, 2019Updated 7 years ago
- Emotion recognition of Speaker's Speech Data. Employ speaker detection classifiers for emotion recognition, a multiclass classification p…☆16Jun 28, 2015Updated 11 years ago
- Speaker diarization scripts, based on AaltoASR☆192Jan 3, 2019Updated 7 years ago
- python wrap for hts engine☆14Jan 30, 2018Updated 8 years ago
- Speaker recognition/identification system in Python. Python3 port.☆14May 2, 2015Updated 11 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆12Jun 13, 2017Updated 9 years ago
- ☆65Dec 20, 2013Updated 12 years ago
- TACOTRON: TOWARDS END-TO-END SPEECH SYNTHESIS☆16Sep 26, 2017Updated 9 years ago
- The Additive Margin MobileNet1D is a new light weight deep learning model for Speaker Recognition which is based on the MobileNetV2 archi…☆31Oct 3, 2023Updated 2 years ago
- Speaker recognition library based on MARF for raspberry pi and other SBCs.☆57Jan 16, 2018Updated 8 years ago
- A simple tutorial on setting up Sparrowhawk - a text-to-speech normalization engine☆14Oct 16, 2017Updated 8 years ago
- A collection of trending speech enhancement papers☆12Dec 4, 2020Updated 5 years ago
- Automatic Speaker Recognition algorithms in Python☆97Sep 25, 2021Updated 5 years ago
- ☆12Feb 9, 2021Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A speaker recognition system which uses GMM-UBM for use in an Android application which helps in monitoring patients suffering from Schiz…☆55Jun 13, 2018Updated 8 years ago
- A GPU language model, based on btree backed tries.☆30Mar 6, 2018Updated 8 years ago
- TristouNet: Triplet Loss for Speaker Turn Embedding☆121Jul 6, 2017Updated 9 years ago
- From a large speech audio file and its corresponding body of text, automatically chunk the audio and text into (phrase, audio_snippet) pa…☆17May 15, 2015Updated 11 years ago
- Extension to Kaldi implementing the standard i-vector hyperparameter estimation and i-vector extraction procedure☆89Feb 23, 2018Updated 8 years ago
- Text-Dependent Speaker Recognition System with Machine Learning Techniques☆10Dec 31, 2017Updated 8 years ago
- Perform the forced decoding with target transcription☆11Sep 12, 2018Updated 8 years ago
- Speaker embedding(verification and recognition) using Tensorflow with Kaldi☆41Sep 18, 2017Updated 9 years ago
- Coqui STT (🐸STT) based forced alignment tool☆13Feb 24, 2022Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Speaker Diarization is the problem of separating speakers in an audio. There could be any number of speakers and final result should stat…☆64Jan 8, 2021Updated 5 years ago
- Character level speech recognizer using ctc loss with deep rnns in TensorFlow.☆78Jun 9, 2018Updated 8 years ago
- Project to explore Speaker and Voice Identification. To follow will be further Speech Recognition tasks.☆52Apr 29, 2019Updated 7 years ago
- Speaker Recognition application using fast-forward NN☆16Jun 14, 2012Updated 14 years ago
- Source code from our RecSys 2020 paper: "Making neural network interpretable with attribution: application to implicit signals prediction…☆14Oct 2, 2020Updated 5 years ago
- Source code for "Unsupervised Lexicon Discovery from Acoustic Input ", Lee et al, 2015 TACL☆10Aug 11, 2016Updated 10 years ago
- Speaker recognition/identification system in Python☆76Sep 20, 2018Updated 8 years ago
- ☆14Dec 7, 2022Updated 3 years ago
- The Additive Margin SincNet (AM-SincNet) is a new approach for speaker recognition problems which is based in the neural network architec…☆46Oct 3, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Speech recognition using Google Cloud Speech API☆25Feb 7, 2017Updated 9 years ago
- Speech Recognition Using Tacotron☆166Sep 20, 2017Updated 9 years ago
- This Repository includes four different implementations of the Speaker Verification task including the GMM_UBM, Ivector, Deep-Speaker, an…☆32Jul 3, 2018Updated 8 years ago
- single channel speech separation for music vocal and accompany separate、voice reduce noise☆14Jul 9, 2019Updated 7 years ago
- Transfer Learning via Unsupervised Task Discovery for Visual Question Answering☆19Apr 8, 2019Updated 7 years ago
- This ist the repository for the term project Speech Recognition using Deep Neural Networks for the course ELEC-E5510-Speech Recognition☆12Dec 8, 2015Updated 10 years ago
- Free noise reduction of speech signals☆12Jul 26, 2016Updated 10 years ago