A family of efficient speech models for multilingual phone recognition
☆72Jul 18, 2026Updated 3 weeks ago
Alternatives and similar repositories for zipa
Users that are interested in zipa are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Keyword spotting and forced alignment in any language☆101Jun 15, 2026Updated last month
- A toolkit and benchmark for evaluating phonetic capabilities of speech models.☆18Apr 10, 2026Updated 4 months ago
- phone inventory library☆17May 15, 2023Updated 3 years ago
- ☆16Mar 19, 2026Updated 4 months ago
- OLaPh (Optimal Language Phonemizer) is a multilingual phonemization framework that converts text into phonemes surpassing the quality of …☆22Jul 20, 2026Updated 3 weeks ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- A universal phone recognizer that can transcribe speech in 70+ languages into IPA☆31Jun 9, 2026Updated 2 months ago
- Dataset of ICASSP 2021 MULTILINGUAL PHONETIC DATASET FOR LOW RESOURCE SPEECH RECOGNITION☆46May 12, 2023Updated 3 years ago
- phoneme tokenizer and grapheme-to-phoneme model for 8k languages☆174Jun 9, 2023Updated 3 years ago
- A Diffrentiable WFST-based End-to-End Automatic Speech Recognition toollkit with flexible topology support☆12Feb 15, 2026Updated 5 months ago
- Pybind11 bindings for Kaldi☆15Jul 11, 2026Updated last month
- Grapheme-to-Phoneme transductions that preserve input and output indices, and support cross-lingual g2p!☆203Jul 24, 2026Updated 2 weeks ago
- pytorch model for contexless-phoneme prediction from speech audio☆32Oct 30, 2025Updated 9 months ago
- Hosts text-to-speech corpus and speech synthesizers for African languages.☆19May 31, 2023Updated 3 years ago
- Building actual open source including dataset Multilingual TTS more than 150 languages with Voice Cloning.☆56Jul 14, 2026Updated 3 weeks ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Repo & Project for the Imminent Research Grant code & tasks☆12May 20, 2024Updated 2 years ago
- Grapheme-to-phoneme tool for corpus conversion, where phonemes match Phoible inventories☆19Apr 10, 2025Updated last year
- asr2k☆51Jun 2, 2024Updated 2 years ago
- A multilingual phoneme recognizer capable of generalizing zero-shot to unseen phoneme inventories.☆30Mar 14, 2025Updated last year
- IPA Phonemizer/Dephonemizer for 140 human languages☆61Jun 20, 2026Updated last month
- This repository contains all the code necessary for running the multilingual distilwhisper from Ferraz et al. 2024 IEEE ICASSP paper.☆35Apr 22, 2026Updated 3 months ago
- PHOIBLE data and development.☆147Updated this week
- A variable-frame-rate 16 kHz speech codec based on FocalCodec☆20Feb 11, 2026Updated 6 months ago
- The EveryVoice TTS Toolkit - Text To Speech for your language☆44Jul 31, 2026Updated last week
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Annotations and scripts for use with University of Wisconsin X-Ray Microbeam Speech Production Database (1994)☆14Oct 8, 2020Updated 5 years ago
- Charsiu: A neural phonetic aligner.☆348Sep 19, 2022Updated 3 years ago
- A collection of utilities for handling IPA phones.☆27Sep 24, 2023Updated 2 years ago
- ☆10Apr 17, 2024Updated 2 years ago
- MaSS - Multilingual corpus of Sentence-aligned Spoken utterances☆50Sep 16, 2024Updated last year
- Dictionary of pairs of Korean word and IPA crawled from Wiktionary (Korean edition)☆23Nov 12, 2025Updated 9 months ago
- NISQA - Non-Intrusive Speech Quality and TTS Naturalness Assessment☆16Apr 13, 2022Updated 4 years ago
- Multilingual G2P in 100 languages☆393May 26, 2023Updated 3 years ago
- IPA tokeniser☆19Jul 28, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- This is the official repository for the HUI-Audio-Corpus-German. The corresponding paper is in the process of publication. With the repo…☆35Mar 31, 2023Updated 3 years ago
- Prosody and Pronunciation Modification Network☆64May 5, 2025Updated last year
- Repository for multilingual speech data resources for native languages of Zambia.☆22Oct 9, 2024Updated last year
- Python package and data files for manipulating phonological segments (phones, phonemes) in terms of universal phonological features.☆321Oct 22, 2025Updated 9 months ago
- On-device VAD / streaming STT / TTS / diarization in C++17 (ONNX + LiteRT) with a voice-agent pipeline. Linux, Windows, Android.☆66Aug 5, 2026Updated last week
- Data and code for grapheme-to-phoneme transducers in lots of languages☆152Apr 5, 2024Updated 2 years ago
- Extract phoneme-level timestamps from speeh audio.☆161Jun 7, 2026Updated 2 months ago