Project to learn about speech recognition - both Speaker Diarization and other Speech Recognition applications.
☆50Feb 1, 2017Updated 9 years ago
Alternatives and similar repositories for Deep-Learning-Speech-Recognition
Users that are interested in Deep-Learning-Speech-Recognition are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Unsupervised Speaker Clustering & Speaker Recognition☆13Jan 7, 2019Updated 7 years ago
- Emotion recognition of Speaker's Speech Data. Employ speaker detection classifiers for emotion recognition, a multiclass classification p…☆16Jun 28, 2015Updated 11 years ago
- Keras implementation of SincNet (https://github.com/mravanelli/SincNet, https://arxiv.org/abs/1808.00158)☆12Aug 5, 2018Updated 7 years ago
- Speaker diarization scripts, based on AaltoASR☆191Jan 3, 2019Updated 7 years ago
- Recognizing a speaker using Deep Learning☆11Dec 25, 2017Updated 8 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- python wrap for hts engine☆14Jan 30, 2018Updated 8 years ago
- Speaker recognition/identification system in Python. Python3 port.☆14May 2, 2015Updated 11 years ago
- Speech Processing & Linguistic Analysis Tool☆11Jun 30, 2019Updated 7 years ago
- ☆12Jun 13, 2017Updated 9 years ago
- TACOTRON: TOWARDS END-TO-END SPEECH SYNTHESIS☆16Sep 26, 2017Updated 8 years ago
- The Additive Margin MobileNet1D is a new light weight deep learning model for Speaker Recognition which is based on the MobileNetV2 archi…☆31Oct 3, 2023Updated 2 years ago
- SQL-based Recommendation System for multi-topic recommendations☆12Jun 1, 2021Updated 5 years ago
- A simple tutorial on setting up Sparrowhawk - a text-to-speech normalization engine☆14Oct 16, 2017Updated 8 years ago
- A collection of trending speech enhancement papers☆11Dec 4, 2020Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Automatic Speaker Recognition algorithms in Python☆96Sep 25, 2021Updated 4 years ago
- ☆12Feb 9, 2021Updated 5 years ago
- A speaker recognition system which uses GMM-UBM for use in an Android application which helps in monitoring patients suffering from Schiz…☆55Jun 13, 2018Updated 8 years ago
- A GPU language model, based on btree backed tries.☆30Mar 6, 2018Updated 8 years ago
- TristouNet: Triplet Loss for Speaker Turn Embedding☆121Jul 6, 2017Updated 9 years ago
- From a large speech audio file and its corresponding body of text, automatically chunk the audio and text into (phrase, audio_snippet) pa…☆17May 15, 2015Updated 11 years ago
- Extension to Kaldi implementing the standard i-vector hyperparameter estimation and i-vector extraction procedure☆88Feb 23, 2018Updated 8 years ago
- Text-Dependent Speaker Recognition System with Machine Learning Techniques☆10Dec 31, 2017Updated 8 years ago
- Speaker recognition and verification with deep learning☆13Mar 7, 2017Updated 9 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Perform the forced decoding with target transcription☆11Sep 12, 2018Updated 7 years ago
- A python implementation of the neural network joint language model and an extension of it using global source context.☆11May 17, 2017Updated 9 years ago
- Speaker embedding(verification and recognition) using Tensorflow with Kaldi☆41Sep 18, 2017Updated 8 years ago
- Coqui STT (🐸STT) based forced alignment tool☆13Feb 24, 2022Updated 4 years ago
- SpeakerVoiceIdentifier can recognize the voice of a speaker by learning.☆35Feb 20, 2017Updated 9 years ago
- Speaker Diarization is the problem of separating speakers in an audio. There could be any number of speakers and final result should stat…☆64Jan 8, 2021Updated 5 years ago
- Character level speech recognizer using ctc loss with deep rnns in TensorFlow.☆78Jun 9, 2018Updated 8 years ago
- Project to explore Speaker and Voice Identification. To follow will be further Speech Recognition tasks.☆52Apr 29, 2019Updated 7 years ago
- Speaker Recognition application using fast-forward NN☆16Jun 14, 2012Updated 14 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Source code for "Unsupervised Lexicon Discovery from Acoustic Input ", Lee et al, 2015 TACL☆10Aug 11, 2016Updated 9 years ago
- Source code from our RecSys 2020 paper: "Making neural network interpretable with attribution: application to implicit signals prediction…☆14Oct 2, 2020Updated 5 years ago
- Speaker recognition/identification system in Python☆76Sep 20, 2018Updated 7 years ago
- ☆14Dec 7, 2022Updated 3 years ago
- Classification of MNIST digits by convolutional neural networks and then extracting features. After that I tune the to classes labels …☆14Dec 31, 2016Updated 9 years ago
- The Additive Margin SincNet (AM-SincNet) is a new approach for speaker recognition problems which is based in the neural network architec…☆46Oct 3, 2023Updated 2 years ago
- System for identifying speaker from given speech signal using MFCC,LPC features and Gaussian Mixture Models☆21Nov 5, 2017Updated 8 years ago