Speech recognition module for Python, supporting several engines and APIs, online and offline.
☆8,987Sep 2, 2026Updated last week
Alternatives and similar repositories for speech_recognition
Users that are interested in speech_recognition are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- DeepSpeech is an open source embedded (offline, on-device) speech-to-text engine which can run in real time on devices ranging from a Ras…☆26,778Jun 19, 2025Updated last year
- 🎙Speech recognition using the tensorflow deep learning framework, sequence-to-sequence neural networks☆2,172Jan 17, 2024Updated 2 years ago
- Python interface to CMU Sphinxbase and Pocketsphinx libraries☆373Jun 27, 2023Updated 3 years ago
- kaldi-asr/kaldi is the official location of the Kaldi project.☆15,475Sep 22, 2025Updated 11 months ago
- A small speech recognizer☆4,340Aug 30, 2026Updated last week
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- End-to-end Automatic Speech Recognition for Madarian and English in Tensorflow☆2,832Mar 24, 2023Updated 3 years ago
- Manipulate audio with a simple and easy high level interface☆9,793Mar 19, 2026Updated 5 months ago
- Speech-to-Text-WaveNet : End-to-end sentence level English speech recognition based on DeepMind's WaveNet and tensorflow☆4,004Oct 8, 2021Updated 4 years ago
- Python Audio Analysis Library: Feature Extraction, Classification, Segmentation and Applications☆6,259Aug 4, 2025Updated last year
- Offline Text To Speech synthesis for python☆2,531Jul 22, 2026Updated last month
- A Deep-Learning-Based Chinese Speech Recognition System 基于深度学习的中文语音识别系统☆8,387Apr 10, 2026Updated 5 months ago
- Python library for audio and music analysis☆8,596Aug 22, 2026Updated 2 weeks ago
- Python library and CLI tool to interface with Google Translate's text-to-speech API☆2,630Apr 6, 2026Updated 5 months ago
- Facebook AI Research's Automatic Speech Recognition Toolkit☆6,437Aug 28, 2026Updated last week
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- End-to-End Speech Processing Toolkit☆9,951Updated this week
- Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node☆15,117Aug 9, 2026Updated last month
- 💫 Industrial-strength Natural Language Processing (NLP) in Python☆33,888Aug 24, 2026Updated 2 weeks ago
- Python module installed with setup.py☆337Jun 29, 2022Updated 4 years ago
- Robust Speech Recognition via Large-Scale Weak Supervision☆108,790Aug 31, 2026Updated last week
- A PyTorch-based Speech Toolkit☆11,811Aug 27, 2026Updated 2 weeks ago
- The world's simplest facial recognition api for Python and the command line☆56,725Jun 25, 2026Updated 2 months ago
- Deep Learning for humans☆64,322Updated this week
- Models and examples built with TensorFlow☆77,658Updated this week
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Speech Recognition using DeepSpeech2.☆2,136Dec 13, 2022Updated 3 years ago
- Future versions with model training module will be maintained through a forked version here: https://github.com/seasalt-ai/snowboy☆3,365Oct 13, 2021Updated 4 years ago
- ChatterBot is a machine learning, conversational dialog engine for creating chat bots☆14,509Aug 25, 2026Updated 2 weeks ago
- This library provides common speech features for ASR including MFCCs and filterbank energies.☆2,423Oct 20, 2021Updated 4 years ago
- Real-time full-duplex speech recognition server, based on the Kaldi toolkit and the GStreamer framwork.☆1,094Jun 8, 2024Updated 2 years ago
- Face recognition with deep neural networks.☆15,439Jul 18, 2026Updated last month
- Neural building blocks for speaker diarization: speech activity detection, speaker change detection, overlapped speech detection, speaker…☆10,525Updated this week
- Python interface to the WebRTC Voice Activity Detector☆2,499Jul 4, 2024Updated 2 years ago
- 🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production☆45,994Aug 16, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- 💬 Open source machine learning framework to automate text- and voice-based conversations: NLU, dialogue management, connect to Slack, …☆21,320Jul 24, 2026Updated last month
- Topic Modelling for Humans☆16,481Nov 1, 2025Updated 10 months ago
- Library for fast text representation and classification.☆26,529Mar 22, 2024Updated 2 years ago
- An Open Source Machine Learning Framework for Everyone☆199,342Updated this week
- Video editing with Python☆14,890Aug 26, 2026Updated 2 weeks ago
- Automatic Speech Recognition (ASR), Speaker Verification, Speech Synthesis, Text-to-Speech (TTS), Language Modelling, Singing Voice Synth…☆3,130Oct 19, 2023Updated 2 years ago
- Simple, Pythonic, text processing--Sentiment analysis, part-of-speech tagging, noun phrase extraction, translation, and more.☆9,546Updated this week