Zoom Audio Transcription offline
☆34Sep 30, 2020Updated 5 years ago
Alternatives and similar repositories for zoom_audio_transcribe
Users that are interested in zoom_audio_transcribe are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Artie Bias Corpus: an audio corpus + code for detecting demographic bias☆20Jul 21, 2020Updated 6 years ago
- 📖 LanMIT: A Toolkit for Improving Language Models in Low-resourced Speech Recognition based on Kaldi.☆22Jul 12, 2019Updated 7 years ago
- Implementation of different noise embeddings for noise aware training of Kaldi acoustic models.☆13Feb 13, 2021Updated 5 years ago
- Using YouTube to prepare a speech recognition dataset for any language☆10Mar 30, 2021Updated 5 years ago
- This repo contains the baseline model recipes and pre-trained model for GramVanni hindi ASR challenge☆16Mar 26, 2022Updated 4 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- CMU multilingual speech repository☆30Apr 15, 2022Updated 4 years ago
- sherpa with mlx☆15Aug 2, 2025Updated last year
- [ICLR 2022] "Audio Lottery: Speech Recognition Made Ultra-Lightweight, Noise-Robust, and Transferable", by Shaojin Ding, Tianlong Chen, Z…☆32Apr 8, 2022Updated 4 years ago
- Java Bindings for the C++ library DeepSpeech☆10Jun 4, 2020Updated 6 years ago
- Unsupervised speech activity detection system.☆11Jul 2, 2018Updated 8 years ago
- ☆14Jun 12, 2015Updated 11 years ago
- Benchmarking different VAD models on AVA-Speech dataset☆19May 21, 2023Updated 3 years ago
- pyMUSHRA is a python web application which hosts webMUSHRA experiments and collects the data with python.☆47Jul 3, 2026Updated last month
- Web page for ISCA Special Interest Group: Robust Speech Processing (RoSP)☆11Dec 4, 2023Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆17Apr 14, 2023Updated 3 years ago
- Simple Kaldi model server for chain (nnet3) models in online recognition mode directly from a local microphone☆35Feb 18, 2022Updated 4 years ago
- an tutorial implement of voice conversion using pytorch☆34Mar 30, 2018Updated 8 years ago
- This repository provides data and code for "Vox Populi, Vox DIY: Benchmark Dataset for Crowdsourced Audio Transcription" paper.☆16Jul 22, 2021Updated 5 years ago
- Automatic Speech Recognition (ASR) system for the Samrómur speech corpus using Kaldi☆12Sep 30, 2022Updated 3 years ago
- Sequence-to-sequence TTS based on Kyubyong's dc_tts☆61Feb 2, 2023Updated 3 years ago
- Multistream CNN for Robust Acoustic Modeling☆40Jun 17, 2021Updated 5 years ago
- ☆22Jul 22, 2022Updated 4 years ago
- Example workflow for our data-centric speech benchmark☆17Jul 6, 2023Updated 3 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Voice conversion training with 109 speakers with limited training samples☆35Dec 21, 2020Updated 5 years ago
- Vim Speech Recognition Experiments☆20May 30, 2025Updated last year
- MaSS - Multilingual corpus of Sentence-aligned Spoken utterances☆50Sep 16, 2024Updated last year
- ☆22Sep 24, 2018Updated 7 years ago
- Analyzes signal, finds fundamental frequency, HNR etc☆15Aug 23, 2017Updated 9 years ago
- Scripts for training Kaldi for German speech recognition (ASR).☆27Feb 11, 2021Updated 5 years ago
- ☆35Nov 24, 2024Updated last year
- ☆38Sep 20, 2022Updated 3 years ago
- Interface for Controllable Expressive Talking Machine☆40Sep 20, 2025Updated 11 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Evaluation of STT models for german language☆16Jan 22, 2022Updated 4 years ago
- Efficient Personalized Speech Enhancement through Self-Supervised Learning☆24Mar 12, 2023Updated 3 years ago
- Detect emotion from audio☆14Nov 20, 2018Updated 7 years ago
- REST api for mozilla deepspeech voice recognition engine☆20Nov 1, 2021Updated 4 years ago
- Neural network density models for speech separation.☆20Nov 26, 2020Updated 5 years ago
- speech-aligner,是一个从“人声语音”及其“语言文本”,产生音素级别时间对齐标注的工具。speech-aligner, is a tool that generate phoneme-level alignment between human speech an…☆15Dec 19, 2018Updated 7 years ago
- A GUI automation tool to export an Ableton Live set with the send effects printed to each stem.☆15Feb 28, 2017Updated 9 years ago