Project to learn about speech recognition - both Speaker Diarization and other Speech Recognition applications.
☆50Feb 1, 2017Updated 9 years ago
Alternatives and similar repositories for Deep-Learning-Speech-Recognition
Users that are interested in Deep-Learning-Speech-Recognition are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Unsupervised Speaker Clustering & Speaker Recognition☆13Jan 7, 2019Updated 7 years ago
- Emotion recognition of Speaker's Speech Data. Employ speaker detection classifiers for emotion recognition, a multiclass classification p…☆16Jun 28, 2015Updated 11 years ago
- Speaker diarization scripts, based on AaltoASR☆192Jan 3, 2019Updated 7 years ago
- Recognizing a speaker using Deep Learning☆11Dec 25, 2017Updated 8 years ago
- python wrap for hts engine☆14Jan 30, 2018Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Speaker recognition/identification system in Python. Python3 port.☆14May 2, 2015Updated 11 years ago
- Speech Processing & Linguistic Analysis Tool☆11Jun 30, 2019Updated 7 years ago
- ☆12Jun 13, 2017Updated 9 years ago
- ☆65Dec 20, 2013Updated 12 years ago
- TACOTRON: TOWARDS END-TO-END SPEECH SYNTHESIS☆16Sep 26, 2017Updated 8 years ago
- The Additive Margin MobileNet1D is a new light weight deep learning model for Speaker Recognition which is based on the MobileNetV2 archi…☆31Oct 3, 2023Updated 2 years ago
- Speaker recognition library based on MARF for raspberry pi and other SBCs.☆57Jan 16, 2018Updated 8 years ago
- SQL-based Recommendation System for multi-topic recommendations☆12Jun 1, 2021Updated 5 years ago
- A simple tutorial on setting up Sparrowhawk - a text-to-speech normalization engine☆14Oct 16, 2017Updated 8 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Train a LSTM neural networks on Vox Forge public audio data set to recognize speaker's gender☆13Mar 26, 2026Updated 5 months ago
- A collection of trending speech enhancement papers☆12Dec 4, 2020Updated 5 years ago
- Automatic Speaker Recognition algorithms in Python☆97Sep 25, 2021Updated 4 years ago
- ☆12Feb 9, 2021Updated 5 years ago
- A GPU language model, based on btree backed tries.☆30Mar 6, 2018Updated 8 years ago
- TristouNet: Triplet Loss for Speaker Turn Embedding☆121Jul 6, 2017Updated 9 years ago
- From a large speech audio file and its corresponding body of text, automatically chunk the audio and text into (phrase, audio_snippet) pa…☆17May 15, 2015Updated 11 years ago
- Extension to Kaldi implementing the standard i-vector hyperparameter estimation and i-vector extraction procedure☆89Feb 23, 2018Updated 8 years ago
- Perform the forced decoding with target transcription☆11Sep 12, 2018Updated 7 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- A python implementation of the neural network joint language model and an extension of it using global source context.☆11May 17, 2017Updated 9 years ago
- Speaker embedding(verification and recognition) using Tensorflow with Kaldi☆41Sep 18, 2017Updated 8 years ago
- Coqui STT (🐸STT) based forced alignment tool☆13Feb 24, 2022Updated 4 years ago
- SpeakerVoiceIdentifier can recognize the voice of a speaker by learning.☆35Feb 20, 2017Updated 9 years ago
- Speaker Diarization is the problem of separating speakers in an audio. There could be any number of speakers and final result should stat…☆64Jan 8, 2021Updated 5 years ago
- Character level speech recognizer using ctc loss with deep rnns in TensorFlow.☆78Jun 9, 2018Updated 8 years ago
- Project to explore Speaker and Voice Identification. To follow will be further Speech Recognition tasks.☆52Apr 29, 2019Updated 7 years ago
- Speaker Recognition application using fast-forward NN☆16Jun 14, 2012Updated 14 years ago
- Source code for "Unsupervised Lexicon Discovery from Acoustic Input ", Lee et al, 2015 TACL☆10Aug 11, 2016Updated 10 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Source code from our RecSys 2020 paper: "Making neural network interpretable with attribution: application to implicit signals prediction…☆14Oct 2, 2020Updated 5 years ago
- Speaker recognition/identification system in Python☆76Sep 20, 2018Updated 7 years ago
- ☆14Dec 7, 2022Updated 3 years ago
- Classification of MNIST digits by convolutional neural networks and then extracting features. After that I tune the to classes labels …☆14Dec 31, 2016Updated 9 years ago
- The Additive Margin SincNet (AM-SincNet) is a new approach for speaker recognition problems which is based in the neural network architec…☆46Oct 3, 2023Updated 2 years ago
- System for identifying speaker from given speech signal using MFCC,LPC features and Gaussian Mixture Models☆21Nov 5, 2017Updated 8 years ago
- Speech recognition using Google Cloud Speech API☆25Feb 7, 2017Updated 9 years ago