Grapheme-to-phoneme tool for corpus conversion, where phonemes match Phoible inventories
☆19Apr 10, 2025Updated last year
Alternatives and similar repositories for g2p-plus
Users that are interested in g2p-plus are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- IPA Phonetic dataset lexicon☆18Jun 20, 2026Updated 2 months ago
- A universal phone recognizer that can transcribe speech in 70+ languages into IPA☆34Jun 9, 2026Updated 2 months ago
- Clean and modernized implementation of FastSpeech2/LightSpeech using IPA☆20Aug 16, 2024Updated 2 years ago
- g2p for english tts☆19Nov 10, 2022Updated 3 years ago
- This is a balanced dataset for English homograph disambiguation (HD), generated with Meta's Llama 2-Chat 70B model.☆22Jan 22, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Export an ONNX graph that performs ISTFT. Designed for TTS models.☆28Apr 23, 2024Updated 2 years ago
- Fast inference engine for DACVAE, a neural audio codec that compresses and reconstructs audio using a convolutional encoder-decoder with …☆21Mar 17, 2026Updated 5 months ago
- A family of efficient speech models for multilingual phone recognition☆77Jul 18, 2026Updated last month
- phone inventory library☆17May 15, 2023Updated 3 years ago
- Accompanying code for paper "Attention-Based Contextual Language Model Adaptation for Speech Recognition", submitted to ACL 2021.☆14Jul 25, 2023Updated 3 years ago
- Text-to-Speech Benchmark☆29Aug 18, 2026Updated last week
- Unofficial Pytorch implementation of SNAC: Speaker-normalized affine coupling layer in flow-based architecture for zero-shot multi-speake…☆57Aug 7, 2023Updated 3 years ago
- Conformer block with Rotary Position Embedding, modified from lucidrains' implement☆19Sep 13, 2024Updated last year
- PolEval 2021 Task 1☆15Jun 28, 2022Updated 4 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- SC-CNN: Effective Speaker Conditioning Method for Zero-Shot Multi-Speaker Text-to-Speech Systems☆39Nov 1, 2023Updated 2 years ago
- [ICASSP 2026] Official code for "Measuring Prosody Diversity in Zero-Shot TTS: A New Metric, Benchmark, and Exploration"☆17Apr 16, 2026Updated 4 months ago
- ☆19Mar 22, 2024Updated 2 years ago
- NVV-SuperBench: Beyond Words, Beyond Quality—Benchmarking Nonverbal Vocalizations in Speech Generation (Interspeech 2026 long paper)☆18Jun 21, 2026Updated 2 months ago
- Vocoder-Free Non-Parallel Conversion of Whispered Speech With Masked Cycle-Consistent Generative Adversarial Networks☆17Aug 18, 2023Updated 3 years ago
- Codebase for 'ParaSpeechCLAP: A Dual-Encoder Speech-Text Model for Rich Stylistic Language-Audio Pretraining'☆25Jun 20, 2026Updated 2 months ago
- 基于vits fastspeech2 visinger的tts模型☆24Mar 9, 2023Updated 3 years ago
- Syllable Segmentation and Cross-Lingual Generalization in a Visually Grounded, Self-Supervised Speech Model☆35Aug 27, 2023Updated 3 years ago
- Sequence to sequence model for Arabic punctuation prediction.☆12Feb 13, 2020Updated 6 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Evaluation of STT models for german language☆16Jan 22, 2022Updated 4 years ago
- PitchVC: Pitch Conditioned Any-to-Many Voice Conversion☆35Jun 6, 2024Updated 2 years ago
- Official Implementation and Dataset of paper - DFADD: The Diffusion and Flow-matching based Audio Deepfake Dataset☆17Apr 7, 2025Updated last year
- NEAL (Nature+Energy Audio Labeller) is an open-source interactive audio data annotation tool.☆20Jul 12, 2026Updated last month
- Extract a target speaker’s clean, non-overlapped speech from multi-speaker audio and export word-safe LJSpeech-style TTS datasets.☆23Jun 14, 2026Updated 2 months ago
- End-To-End SpeechSynthesis system with knowledge distillation☆18Jul 16, 2022Updated 4 years ago
- ☆12Nov 7, 2024Updated last year
- ☆46Oct 24, 2020Updated 5 years ago
- 44100Hz日本語音源に対応させた unofficial vits2-TTS implementation in pytorchです。☆24Sep 1, 2023Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- All-in-one Speech Transcription☆11Jun 5, 2026Updated 2 months ago
- ☆17Oct 24, 2025Updated 10 months ago
- C++ version of pyannote audio overlapped speech detection pipeline☆13Feb 14, 2024Updated 2 years ago
- Glow-TTS with Stochastic Duration Predictor and Stochastic Pitch Predictor☆19Jun 5, 2023Updated 3 years ago
- ☆43Feb 8, 2025Updated last year
- version 4.x of the Princeton Geniza Project☆13Aug 20, 2026Updated last week
- ☆63Sep 18, 2022Updated 3 years ago