VAD(Voice Activity Detector) python 实现对时时读入的流式数据进行端点检测
☆49Apr 8, 2015Updated 11 years ago
Alternatives and similar repositories for pyvad
Users that are interested in pyvad are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A Python interface to OpenFst (fix FstDrawer interface issue for 1.6 version)☆17Apr 2, 2018Updated 8 years ago
- 基于双门限识别的语音端点检测系统☆24Jan 16, 2018Updated 8 years ago
- Pronunciation Evaluation☆104Jul 20, 2025Updated last year
- Detects scene change or cuts in a video file☆11Oct 23, 2017Updated 8 years ago
- ☆25Jun 14, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Text to Speech Synthesis based on controllable latent representation☆14Aug 30, 2019Updated 7 years ago
- speech-dereverberation-using-GANs☆13Jan 28, 2019Updated 7 years ago
- Code for SLT 2016 paper on Grapheme-to-Phoneme conversion using attention based encoder-decoder models☆15Feb 20, 2019Updated 7 years ago
- simple dnn based vad☆69Dec 2, 2018Updated 7 years ago
- System for identifying speaker from given speech signal using MFCC,LPC features and Gaussian Mixture Models☆21Nov 5, 2017Updated 8 years ago
- webrtc_rnnvad☆24Jul 12, 2021Updated 5 years ago
- using microphone☆16Sep 2, 2021Updated 5 years ago
- ToneNet: A CNN Model of Tone Classification of Mandarin Chinese☆20Nov 27, 2019Updated 6 years ago
- Source code for "Unsupervised Lexicon Discovery from Acoustic Input ", Lee et al, 2015 TACL☆10Aug 11, 2016Updated 10 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Voice Activity Detector☆74Mar 7, 2026Updated 6 months ago
- Voice Activity Detector in Python☆481Nov 17, 2020Updated 5 years ago
- Improving the Goodness of Pronunciation with DNNs and RNNs☆32Sep 26, 2018Updated 7 years ago
- An example directory for running Multi-Task Learning training on Kaldi neural networks. In Kaldi-speak, this is an egs dir for nnet3 trai…☆55Jan 2, 2020Updated 6 years ago
- ☆20Nov 22, 2020Updated 5 years ago
- ☆76Mar 18, 2022Updated 4 years ago
- A KALDI/C++ implementation of GoogleBrain's SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition☆15Sep 4, 2019Updated 7 years ago
- python script for voice activity detection.☆36Aug 16, 2024Updated 2 years ago
- Voice agent on edge device☆15Sep 6, 2026Updated 2 weeks ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- voice active detection (python ver/simple and easy-to-use)☆12May 1, 2017Updated 9 years ago
- Goodness of Pronunciation (GOP) for oral reading assessment.☆55Nov 17, 2021Updated 4 years ago
- a Chinese tokenizer☆18Jun 16, 2013Updated 13 years ago
- ☆33Aug 6, 2021Updated 5 years ago
- Meta-embeddings are a probabilistic generalization of embeddings in machine learning.☆23Nov 23, 2018Updated 7 years ago
- A tensorflow implementation of ByteNet with layer masking.☆10Jun 29, 2017Updated 9 years ago
- Code for the paper: Audio to Score Matching by Combining Phonetic and Duration Information☆29Jul 9, 2017Updated 9 years ago
- An open-source tool for automatic speech recognition ASR quality estimation.☆24Dec 12, 2019Updated 6 years ago
- Reproduction of a paper"Small-footprint keyword spotting using deep neural networks"☆12Mar 11, 2019Updated 7 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Pronunciation-assisted Subword Modeling☆31May 30, 2019Updated 7 years ago
- Monitoring the driver's behavior, recording their driving event and providing feedback☆10Jan 29, 2018Updated 8 years ago
- This repository is developed in MATLAB. Speech Augmentation is based on Adaptive Filtering while Endpoint Detection is based on Voice Act…☆10Dec 7, 2020Updated 5 years ago
- 基于机器视觉的智能辅助驾驶应用,作为系统的移动端衍生,使用Flutter进行开发,适配双端(Android、IOS)☆12Jun 16, 2019Updated 7 years ago
- Generate Huffman codes with Python☆20Feb 1, 2022Updated 4 years ago
- Taking Live Camera Feed from Android to Python on Desktop Wirelessly☆12Apr 16, 2017Updated 9 years ago
- CAT is more than a CRF-based ASR toolkit: it provides a complete workflow for data-efficient end-to-end ASR, supporting CTC, CTC-CRF, RNN…☆369Feb 5, 2026Updated 7 months ago