A simple example for use speech recognition baidu api with python.
☆115Apr 8, 2021Updated 5 years ago
Alternatives and similar repositories for python-Speech_Recognition
Users that are interested in python-Speech_Recognition are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- solutions for https://www.kaggle.com/c/tensorflow-speech-recognition-challenge☆31Jan 28, 2018Updated 8 years ago
- audio cfeatures extraction tool from wav to h5features format☆19May 24, 2019Updated 7 years ago
- Realtime sound spectrograph built with python, pyaudio, numpy and opencv☆20Apr 16, 2011Updated 15 years ago
- ☆12Jun 26, 2015Updated 11 years ago
- 语智科技远场(单麦克风)语音识别引擎 FFASR 接入指南☆15Aug 4, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 树莓派语音识别机器人(项目转移到autohome项目)☆228Apr 17, 2017Updated 9 years ago
- speaker recognition using keras☆36Nov 29, 2022Updated 3 years ago
- Paderwasn is a collection of methods for acoustic signal processing in wireless acoustic sensor networks (WASNs).☆20May 8, 2025Updated last year
- 未来杯语音赛道说话人识别的baseline☆49Apr 9, 2019Updated 7 years ago
- Base on MFCC and GMM(基于MFCC和高斯混合模型的语音识别)☆253Mar 13, 2019Updated 7 years ago
- 【中文语音识别 】【验证码识别】☆119Jun 17, 2023Updated 3 years ago
- 根据MFCC提取音频特征,训练“飞鱼秀”音频节目语音和音乐的切割。☆30Dec 28, 2017Updated 8 years ago
- Pepper Robot Enhanced Human Interaction☆14Dec 8, 2022Updated 3 years ago
- Speech Recognition with Python examples☆248Oct 2, 2020Updated 5 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Feedforward Sequential Memory Networks (FSMN) implemented by tensorflow☆52Dec 11, 2016Updated 9 years ago
- ICLR 2019 Paper, "Characterizing Audio Adversarial Examples using Temporal Dependency".☆11Apr 3, 2019Updated 7 years ago
- ☆13Oct 10, 2017Updated 8 years ago
- Voice activity detection (VAD) toolkit including DNN, bDNN, LSTM and ACAM based VAD. We also provide our directly recorded dataset.☆869Jun 9, 2021Updated 5 years ago
- This is a GUI for manipulating audio files in the wavelet domain.☆12Jan 31, 2023Updated 3 years ago
- A Demo of Mandarin/Chinese TTS frontend☆284Apr 18, 2022Updated 4 years ago
- ☆10Sep 18, 2017Updated 8 years ago
- Speech recognition module for Python, supporting several engines and APIs, online and offline.☆8,985Jul 31, 2026Updated 2 weeks ago
- speech-aligner,是一个从“人声语音”及其“语言文本”,产生音素级别时间对齐标注的工具。speech-aligner, is a tool that generate phoneme-level alignment between human speech an…☆15Dec 19, 2018Updated 7 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- How to setup Pepper's wifi when bringing it into a new building☆11Jun 23, 2021Updated 5 years ago
- This project uses Pepper robot for human tracking, RGBD data acquisition and physical interaction with people.☆14Feb 8, 2017Updated 9 years ago
- This repository contains scripts for Human Activity Recognition (HAR) project☆15Jan 23, 2015Updated 11 years ago
- A pytorch implementation of the paper : Acoustic Scene Classification with Multiple Decision Schemes.☆20Dec 12, 2020Updated 5 years ago
- A simple pyaudio microphone interface☆11Jul 27, 2018Updated 8 years ago
- Wake-Up-Word Keyword Spotting implemented in Keras☆35Oct 1, 2017Updated 8 years ago
- Software for Decoding of High Order Ambisonics to Irregular Layouts☆13Mar 20, 2014Updated 12 years ago
- Fine-tune Inception v3 for muli-label classification on HICO dataset in TensorFlow☆24Oct 4, 2017Updated 8 years ago
- Smart Glasses for Police Force, a wearable augmented reality glasses with applications in security, medical and industrial field applicat…☆21Mar 20, 2018Updated 8 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Segmentation algorithm for MIREX 2014☆14Dec 16, 2015Updated 10 years ago
- Targeted Adversarial Examples on Speech-to-Text systems☆11Sep 28, 2020Updated 5 years ago
- A Deep Convolutional Neural Network (DCNN) designed for the task of localizing human speech to 168 location classes using binaural microp…☆10Dec 16, 2017Updated 8 years ago
- Multiobjective Optimization Training of PLDA for Speaker Verification☆10Jun 14, 2018Updated 8 years ago
- Text-Dependent Speaker Recognition System with Machine Learning Techniques☆10Dec 31, 2017Updated 8 years ago
- ☆13Jul 16, 2013Updated 13 years ago
- Ossian: A simple language-independent Text-to-speech frontend☆17Mar 1, 2018Updated 8 years ago