☆10Sep 19, 2018Updated 8 years ago
Alternatives and similar repositories for end-point-detection
Users that are interested in end-point-detection are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Repository for the web pages and scripts associated with OpenSLR: the open speech and language repository☆27Jul 26, 2020Updated 6 years ago
- "Recurrent Models of Visual Attention" in TensorFlow☆41Apr 13, 2017Updated 9 years ago
- steps to perform text-based speaker diarization with kaldi toolkit☆12Nov 2, 2018Updated 7 years ago
- This is an extension of kaldi speech recognition software which allows to perform decoding of speech with hybrid word and phoneme graphs.…☆11Feb 4, 2020Updated 6 years ago
- DNN-based speech enhancement using Tensorflow by Haoyu Li (Tokyo univ.)☆17Aug 31, 2017Updated 9 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- This repository creates speaker diarization recipes to be used within the egs folder of kaldi.☆17Aug 12, 2024Updated 2 years ago
- a optional way to extract audio feature☆14Jun 10, 2017Updated 9 years ago
- [EMNLP 2024 Tutorial] Language Agents: Foundations, Prospects, and Risks☆10Nov 27, 2024Updated last year
- Deep neural network based speech enhancement toolkit☆220Jun 14, 2019Updated 7 years ago
- A KALDI/C++ implementation of GoogleBrain's SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition☆15Sep 4, 2019Updated 7 years ago
- Auto-KWS 2021 Challenge 1st place solution.☆11Jul 20, 2021Updated 5 years ago
- Pytorch implementation of 'Improving Self-supervised Lightweight Model Learning via Hard-aware Metric Distillation. In ECCV 2022'☆11Mar 22, 2023Updated 3 years ago
- python c-module for siphash☆20Updated this week
- A packaged convolutional voice activity detector for noisy environments.☆14Jun 15, 2019Updated 7 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆13Mar 25, 2021Updated 5 years ago
- Easier analysis of large speech corpora☆25Jun 22, 2021Updated 5 years ago
- VAD + resampling | High resolution spectrogram☆14Nov 29, 2022Updated 3 years ago
- PyTorch implementation of "Jasper: An End-to-End Convolutional Neural Acoustic Model" (INTERSPEECH 2019)☆32Mar 4, 2021Updated 5 years ago
- This repository☆32Nov 13, 2022Updated 3 years ago
- Convert Numerical Representations to Korean Pronunciation☆14Apr 20, 2020Updated 6 years ago
- Mel-Generalized Cepstrum analysis☆19Jul 21, 2017Updated 9 years ago
- This repository includes the code to reproduce our paper Partially-Connected Differentiable Architecture Search for Deepfake and Spoofing…☆18Apr 30, 2022Updated 4 years ago
- Construct GMM-HMM and Implement the Viterbi algorithm for continuous speech recognition☆15Apr 1, 2018Updated 8 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Spoken Language Identification from Short Utterances☆13Jul 6, 2022Updated 4 years ago
- ☆15Feb 25, 2023Updated 3 years ago
- Script to simulate room impulse responses