MaSS - Multilingual corpus of Sentence-aligned Spoken utterances
☆50Sep 16, 2024Updated last year
Alternatives and similar repositories for mass-dataset
Users that are interested in mass-dataset are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆20Apr 5, 2021Updated 5 years ago
- scipts for working with open.bible data☆26Jan 24, 2022Updated 4 years ago
- A library of speech gadgets.☆15Oct 15, 2022Updated 3 years ago
- ☆14Jun 12, 2015Updated 11 years ago
- Using YouTube to prepare a speech recognition dataset for any language☆10Mar 30, 2021Updated 5 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- unsupervised ASR (mainly phone classifier) using EODM and GAN☆12Oct 22, 2020Updated 5 years ago
- Convert words to numbers☆21Apr 13, 2022Updated 4 years ago
- ☆23Apr 8, 2022Updated 4 years ago
- A curated list of research papers and resources on Indonesian languages☆41Mar 21, 2024Updated 2 years ago
- Google's TPGST reimplementation.☆34Dec 11, 2019Updated 6 years ago
- 🫠 check your data, before you wreck your model☆16Aug 11, 2022Updated 4 years ago
- This repo contains the baseline model recipes and pre-trained model for GramVanni hindi ASR challenge☆16Mar 26, 2022Updated 4 years ago
- ☆17Nov 25, 2019Updated 6 years ago
- phone inventory library☆17May 15, 2023Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Python wrapper for phonetisaurus grapheme to phoneme tool☆12Mar 11, 2021Updated 5 years ago
- Gamma Agreement in Python☆46Mar 4, 2024Updated 2 years ago
- End-to-end Speech Translation☆35Apr 12, 2021Updated 5 years ago
- A repo listing known open source voice tools, ordered by where they sit in the voice stack☆28Sep 23, 2022Updated 3 years ago
- 📖 LanMIT: A Toolkit for Improving Language Models in Low-resourced Speech Recognition based on Kaldi.☆22Jul 12, 2019Updated 7 years ago
- Scripts to create speech corpora from open.bible☆13Jan 3, 2022Updated 4 years ago
- A family of efficient speech models for multilingual phone recognition☆74Jul 18, 2026Updated 3 weeks ago
- python wrap for hts engine☆14Jan 30, 2018Updated 8 years ago
- Neural model for prediction of stress position in Russian words☆13Jun 22, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Bayesian spEEch Recognizer☆55Jan 11, 2021Updated 5 years ago
- ☆11Sep 5, 2025Updated 11 months ago
- Multilingual grapheme-to-phoneme conversion☆20Feb 23, 2018Updated 8 years ago
- NISQA - Non-Intrusive Speech Quality and TTS Naturalness Assessment☆16Apr 13, 2022Updated 4 years ago
- SIGMORPHON 2020 Shared Task: Grapheme-to-Phoneme, Unsupervised Induction of Morphology, and Typologically Diverse Morphological Inflectio…☆36Apr 25, 2025Updated last year
- A python tool that converts Arabic diacritised text to a sequence of phonemes and creates a pronunciation dictionary. This code is based …☆15Sep 5, 2017Updated 8 years ago
- A handy dataset of noises for ASR☆22May 29, 2019Updated 7 years ago
- Web page for ISCA Special Interest Group: Robust Speech Processing (RoSP)☆11Dec 4, 2023Updated 2 years ago
- wake-up word emotion recognition [APSIPA 2022]☆17Nov 11, 2022Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Linguistic processing for Common Voice☆59Jan 18, 2024Updated 2 years ago
- A repository for dictionaries to be used with the Prosodylab-Aligner☆17May 13, 2014Updated 12 years ago
- Vocal Tract Area Estimation by Gradient Descent☆39Jul 16, 2023Updated 3 years ago
- The YouTube Text-To-Speech dataset is comprised of waveform audio extracted from YouTube videos alongside their English transcriptions☆53Apr 1, 2021Updated 5 years ago
- ☆22Jul 22, 2022Updated 4 years ago
- Simple Kaldi recipe for forced alignment☆11Jul 16, 2023Updated 3 years ago
- PyTorch implementation of Retriever: Learning Content-Style Representation☆12Jan 27, 2023Updated 3 years ago