Zero-Resource Speech Discovery, Search, and Evaluation Tools
☆29Aug 6, 2015Updated 11 years ago
Alternatives and similar repositories for ZRTools
Users that are interested in ZRTools are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Software to apply unsupervised word segmentation on lattices or text sequences using a nested hierarchical Pitman Yor language model☆17Nov 24, 2016Updated 9 years ago
- Taiwanese Speech Synthesis with Tacotron2☆26Oct 2, 2022Updated 3 years ago
- Correspondence and autoencoder neural network training for speech using Pylearn2.☆14Dec 9, 2015Updated 10 years ago
- All you need to get started for the Zero Speech Challenge 2017☆47Apr 23, 2019Updated 7 years ago
- Understanding and Tackling Hallucinations in Large Audio-Language Models | ICASSP 2025, Interspeech 2024☆34Mar 14, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆32Jul 27, 2020Updated 6 years ago
- Phonetic and phonological vocoding platform☆17Nov 23, 2016Updated 9 years ago
- This will hold the crowdsourcing platform to be used to store voice data from various speakers which will act as input dataset for speech…☆17Mar 6, 2023Updated 3 years ago
- ASR library☆14Dec 3, 2018Updated 7 years ago
- Prosodic features for machine-learning applications, in Matlab.☆15Oct 14, 2025Updated 10 months ago
- ☆31Jul 13, 2023Updated 3 years ago
- ☆12Feb 26, 2018Updated 8 years ago
- ☆76Mar 18, 2022Updated 4 years ago
- A python implementation of the neural network joint language model and an extension of it using global source context.☆11May 17, 2017Updated 9 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Learnable STRF, from Riad et al. 2021 JASA☆13Aug 21, 2021Updated 5 years ago
- Datasets and Transforms specific to ASR☆17Feb 16, 2017Updated 9 years ago
- A GPU language model, based on btree backed tries.☆30Mar 6, 2018Updated 8 years ago
- Software for unsupervised word segmentation and language model learning using lattices☆46Aug 17, 2016Updated 10 years ago
- The Additive Margin MobileNet1D is a new light weight deep learning model for Speaker Recognition which is based on the MobileNetV2 archi…☆31Oct 3, 2023Updated 2 years ago
- An Empirical Comparison of Unsupervised Constituency Parsing Methods☆14Aug 15, 2021Updated 5 years ago
- MobileNet trained with VoxCeleb dataset and used for voice verification☆18Oct 26, 2022Updated 3 years ago
- Korean speech recognition based on transformer (트랜스포머 기반 한국어 음성 인식)☆31Feb 19, 2021Updated 5 years ago
- Unsupervised Voice Activity Detection by Modeling Source and System Information using Zero Frequency Filtering☆23Oct 19, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Unsupervised speech activity detection system.☆11Jul 2, 2018Updated 8 years ago
- Wenet speech to text for react native☆10Nov 1, 2022Updated 3 years ago
- This is an extension of kaldi speech recognition software which allows to perform decoding of speech with hybrid word and phoneme graphs.…☆11Feb 4, 2020Updated 6 years ago
- Humphrey, E. J. "An Exploration of Deep Learning in Music Informatics." (2015) New York University.☆14Feb 23, 2016Updated 10 years ago
- voice active detection (python ver/simple and easy-to-use)☆12May 1, 2017Updated 9 years ago
- A simple tutorial on setting up Sparrowhawk - a text-to-speech normalization engine☆14Oct 16, 2017Updated 8 years ago
- A transcription text editor with respeak module☆14Jan 24, 2026Updated 7 months ago
- REBORN: Reinforcement-Learned Boundary Segmentation with Iterative Training for Unsupervised ASR☆16Dec 11, 2024Updated last year
- Convert words to numbers☆21Apr 13, 2022Updated 4 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆13Sep 25, 2024Updated last year
- Dynamic time warping (DTW) functions for specifically speech alignment.☆30May 6, 2024Updated 2 years ago
- ☆43Nov 18, 2025Updated 9 months ago
- ☆19May 16, 2015Updated 11 years ago
- ☆10Sep 19, 2022Updated 3 years ago
- steps to perform text-based speaker diarization with kaldi toolkit☆12Nov 2, 2018Updated 7 years ago
- Repository for reproducing result in journal "Self-supervised learning for Speech Emotion Recognition"☆10Mar 15, 2023Updated 3 years ago