A universal phone recognizer that can transcribe speech in 70+ languages into IPA
☆30Jun 9, 2026Updated last month
Alternatives and similar repositories for PhoneticXeus
Users that are interested in PhoneticXeus are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A toolkit and benchmark for evaluating phonetic capabilities of speech models.☆18Apr 10, 2026Updated 3 months ago
- [ICASSP 2026] Official code for "Measuring Prosody Diversity in Zero-Shot TTS: A New Metric, Benchmark, and Exploration"☆17Apr 16, 2026Updated 3 months ago
- Keyword spotting and forced alignment in any language☆101Jun 15, 2026Updated last month
- A family of efficient speech models for multilingual phone recognition☆70Jul 18, 2026Updated 2 weeks ago
- Grapheme-to-phoneme tool for corpus conversion, where phonemes match Phoible inventories☆19Apr 10, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Building actual open source including dataset Multilingual TTS more than 150 languages with Voice Cloning.☆56Jul 14, 2026Updated 2 weeks ago
- SynTTS-Commands is a large-scale, multilingual (English & Chinese) synthetic speech command dataset designed for low-power Keyword Spotti…☆17Feb 5, 2026Updated 5 months ago
- Syllable Segmentation and Cross-Lingual Generalization in a Visually Grounded, Self-Supervised Speech Model☆35Aug 27, 2023Updated 2 years ago
- Koel Labs innovates open-source speech research, inclusive speech technologies, and real-time pronunciation feedback for language learner…☆25Jul 13, 2026Updated 3 weeks ago
- IPA Phonetic dataset lexicon☆18Jun 20, 2026Updated last month
- This is not remotely close to a finished product, and does not intend to nor does this claim to be working fine-tuning code for MaskGCT. …☆13Dec 4, 2024Updated last year
- Code for CTC-TTS: LLM-based dual-streaming text-to-speech with CTC alignment, Interspeech 2026.☆20Jun 9, 2026Updated last month
- OLaPh (Optimal Language Phonemizer) is a multilingual phonemization framework that converts text into phonemes surpassing the quality of …☆20Jul 20, 2026Updated 2 weeks ago
- Pocket TTS but pure C implementation inspired by Flux2.c☆23Feb 14, 2026Updated 5 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ru-normalizr — лучший нормализатор русского текста без LLM. Приводит числа, даты, время, сокращения, римские цифры, символы и латиницу в …☆19Updated this week
- Arabic Grapheme-to-Phoneme (G2P) Conversion☆15Mar 15, 2025Updated last year
- A general purpose task-agnostic speech augmentation policy☆16Mar 13, 2026Updated 4 months ago
- ☆19Aug 27, 2018Updated 7 years ago
- ARCH: Audio Representations benCHmark☆57Aug 26, 2024Updated last year
- Fastest Open Source TTS Model☆78Jul 25, 2026Updated last week
- Conformer block with Rotary Position Embedding, modified from lucidrains' implement☆19Sep 13, 2024Updated last year
- Extract a target speaker’s clean, non-overlapped speech from multi-speaker audio and export word-safe LJSpeech-style TTS datasets.☆21Jun 14, 2026Updated last month
- Estonian text-to-speech text normalization pipeline☆14Dec 17, 2025Updated 7 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆12Nov 7, 2024Updated last year
- IPA Phonemizer/Dephonemizer for 140 human languages☆61Jun 20, 2026Updated last month
- ☆16Mar 19, 2026Updated 4 months ago
- All-in-one Speech Transcription☆11Jun 5, 2026Updated last month
- This is the code of the ICASSP 2020 paper "Joint phoneme alignment and text-informed speech separation on highly corrupted speech"☆16Apr 8, 2024Updated 2 years ago
- ☆17Nov 17, 2020Updated 5 years ago
- A multilingual phoneme recognizer capable of generalizing zero-shot to unseen phoneme inventories.☆30Mar 14, 2025Updated last year
- FestPB é um projeto com objetivo de oferecer suporte ao Português Brasileiro ao software Text-to-Speech Festival Speech Synthesis. Com op…☆10May 5, 2024Updated 2 years ago
- Word alignments based on character-level attention maps in Whisper with unsupervised head selection.☆20Jan 15, 2026Updated 6 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- An upgrade framework for train and validate compare with icefall using Lightning.☆16Mar 26, 2025Updated last year
- Keyword Spotting using BCResNet and Arcface Loss☆13Jan 28, 2022Updated 4 years ago
- Fast linear discrete time filtering in PyTorch.☆32Jul 18, 2026Updated 2 weeks ago
- Project for HIDING SPEAKER’S SEX IN SPEECH USING ZERO-EVIDENCE SPEAKER REPRESENTATION IN AN ANALYSIS/SYNTHESIS PIPELINE☆15Nov 30, 2022Updated 3 years ago
- Experiments with the Mojo 🔥 programming language on macOS arm64 guided by tests☆14Jan 8, 2026Updated 6 months ago
- Free and Open Platform for AI-assisted Computing☆10May 19, 2019Updated 7 years ago
- Contains the code associated with the ICLR submission for our text-to-speech diffusion model☆57Oct 31, 2023Updated 2 years ago