VAD(Voice Activity Detector) python 实现对时时读入的流式数据进行端点检测
☆49Apr 8, 2015Updated 11 years ago
Alternatives and similar repositories for pyvad
Users that are interested in pyvad are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A Python interface to OpenFst (fix FstDrawer interface issue for 1.6 version)☆17Apr 2, 2018Updated 8 years ago
- 基于双门限识别的语音端点检测系统☆24Jan 16, 2018Updated 8 years ago
- Pronunciation Evaluation☆101Jul 20, 2025Updated last year
- Detects scene change or cuts in a video file☆11Oct 23, 2017Updated 8 years ago
- Text to Speech Synthesis based on controllable latent representation☆14Aug 30, 2019Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- speech-dereverberation-using-GANs☆13Jan 28, 2019Updated 7 years ago
- simple dnn based vad☆69Dec 2, 2018Updated 7 years ago
- System for identifying speaker from given speech signal using MFCC,LPC features and Gaussian Mixture Models☆21Nov 5, 2017Updated 8 years ago
- webrtc_rnnvad☆24Jul 12, 2021Updated 5 years ago
- ToneNet: A CNN Model of Tone Classification of Mandarin Chinese☆20Nov 27, 2019Updated 6 years ago
- Source code for "Unsupervised Lexicon Discovery from Acoustic Input ", Lee et al, 2015 TACL☆10Aug 11, 2016Updated 9 years ago
- Voice Activity Detector☆74Mar 7, 2026Updated 4 months ago
- Voice Activity Detector in Python☆481Nov 17, 2020Updated 5 years ago
- Improving the Goodness of Pronunciation with DNNs and RNNs☆32Sep 26, 2018Updated 7 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- An example directory for running Multi-Task Learning training on Kaldi neural networks. In Kaldi-speak, this is an egs dir for nnet3 trai…☆55Jan 2, 2020Updated 6 years ago
- ☆20Nov 22, 2020Updated 5 years ago
- ☆76Mar 18, 2022Updated 4 years ago
- A KALDI/C++ implementation of GoogleBrain's SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition☆15Sep 4, 2019Updated 6 years ago
- python script for voice activity detection.☆36Aug 16, 2024Updated last year
- voice active detection (python ver/simple and easy-to-use)☆12May 1, 2017Updated 9 years ago
- Goodness of Pronunciation (GOP) for oral reading assessment.☆55Nov 17, 2021Updated 4 years ago
- a Chinese tokenizer☆18Jun 16, 2013Updated 13 years ago
- use electron-vue implements springboot's code-generator☆13Jan 13, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆33Aug 6, 2021Updated 4 years ago
- Meta-embeddings are a probabilistic generalization of embeddings in machine learning.☆23Nov 23, 2018Updated 7 years ago
- A tensorflow implementation of ByteNet with layer masking.☆10Jun 29, 2017Updated 9 years ago
- Code for the paper: Audio to Score Matching by Combining Phonetic and Duration Information☆29Jul 9, 2017Updated 9 years ago
- An open-source tool for automatic speech recognition ASR quality estimation.☆24Dec 12, 2019Updated 6 years ago
- Query tuple info in relation file and clean hint info in tuple, change transaction status in commit log file.☆10Dec 6, 2015Updated 10 years ago
- Reproduction of a paper"Small-footprint keyword spotting using deep neural networks"☆12Mar 11, 2019Updated 7 years ago
- Pronunciation-assisted Subword Modeling☆31May 30, 2019Updated 7 years ago
- This repository is developed in MATLAB. Speech Augmentation is based on Adaptive Filtering while Endpoint Detection is based on Voice Act…☆10Dec 7, 2020Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- The goal of this project was to build a smartphone application that can recognize hazardous behavior of a car driver. The driver actions …☆10Nov 15, 2019Updated 6 years ago
- Generate Huffman codes with Python☆20Feb 1, 2022Updated 4 years ago
- Taking Live Camera Feed from Android to Python on Desktop Wirelessly☆12Apr 16, 2017Updated 9 years ago
- CAT is more than a CRF-based ASR toolkit: it provides a complete workflow for data-efficient end-to-end ASR, supporting CTC, CTC-CRF, RNN…☆368Feb 5, 2026Updated 5 months ago
- ☆11Apr 10, 2021Updated 5 years ago
- This project senses human activity with the help of an app installed in the smartphone which uses ML to predict from the data collected f…☆14Feb 13, 2021Updated 5 years ago
- Demo for Neural Spatio-Temporal Beamformer for Target Speech Separation accepted to INTERSPEECH2020☆16Oct 20, 2020Updated 5 years ago