Montreal Forced Aligner for Vietnamese
☆15Oct 23, 2023Updated 2 years ago
Alternatives and similar repositories for ViMFA
Users that are interested in ViMFA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- LIGHTVOC AN UPSAMPLING-FREE GAN VOCODER BASED ON CONFORMER AND INVERSE SHORT-TIME FOURIER TRANSFORM☆18May 17, 2024Updated 2 years ago
- Vi_G2P or ViG2P: G2P package for Vietnamese: based on vPhon and phonology knowledge to convert Raw text - Graphoneme to IPA☆110Jun 21, 2024Updated 2 years ago
- Python - NSW package for Vietnamese: Normalization system to convert numbers, abbreviations, and words that cannot be pronounced into syl…☆67Jan 1, 2025Updated last year
- Simple text to phonemes converter for multiple languages☆20Nov 21, 2022Updated 3 years ago
- ☆15Jul 24, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Generate degraded speech datasets for noise-robust ASR benchmarking☆46Jun 9, 2026Updated 4 months ago
- An official implementation of Style-Talker for Spoken Dialogue Generation☆23Jan 12, 2025Updated last year
- A Vietnamese phonetizer☆56May 29, 2024Updated 2 years ago
- Vietnamese Voice Cloning System using Speaker Verification training on multispeaker VITS☆55Dec 1, 2023Updated 2 years ago
- StyleTTS-ZS: Efficient High-Quality Zero-Shot Text-to-Speech Synthesis with Distilled Time-Varying Style Diffusion☆10Sep 22, 2024Updated 2 years ago
- Sing any popular song with your voice☆11Jul 10, 2022Updated 4 years ago
- Dictionary of pairs of Korean word and IPA crawled from Wiktionary (Korean edition)☆24Nov 12, 2025Updated 10 months ago
- Once more Diarization: Improving meeting transcription systems through segment-level speaker reassignment☆14Feb 5, 2025Updated last year
- A repo containing download guidance and corresponding scripts of the VoxBlink dataset.☆30Apr 16, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆38Oct 10, 2025Updated last year
- End-to-End Chinese Speaker Identification☆12Nov 17, 2022Updated 3 years ago
- Text frontend for ESPnet tts recipes☆35Sep 2, 2026Updated last month
- This repository contains source codes for SoftCTC. Original paper can be found here: https://arxiv.org/abs/2212.02135☆19Mar 7, 2023Updated 3 years ago
- A Vietnamese Text-to-Speech library that provides high-quality speech synthesis with voice cloning capabilities☆103Jul 14, 2025Updated last year
- [ACL 2025] OZSpeech: One-step Zero-shot Speech Synthesis with Learned-Prior-Conditioned Flow Matching☆45Feb 9, 2025Updated last year
- Code for "Weakly-supervised Fingerspelling Recognition in British Sign Language Videos", BMVC 2022.☆12Jun 22, 2023Updated 3 years ago
- ☆56Sep 3, 2025Updated last year
- A synthesized dataset for Vietnamese TTS task☆66May 6, 2022Updated 4 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- ☆13Apr 12, 2025Updated last year
- Dự án công cụ chuyển đổi giọng nói dành cho người Việt☆33Sep 21, 2026Updated 2 weeks ago
- [ICASSP'23] Online speaker clustering☆19Feb 22, 2026Updated 7 months ago
- ViSen is library to format tone of Vietnamese sentences☆22Nov 9, 2021Updated 4 years ago
- sherpa with mlx☆15Aug 2, 2025Updated last year
- ☆10Nov 27, 2024Updated last year
- ViStreamASR - Real-Time Vietnamese Speech Recognition☆63Jul 12, 2025Updated last year
- A library for multilingual word, phrase and sentence segmentation.☆17Sep 2, 2026Updated last month
- Handling audio files in Python☆39Sep 15, 2026Updated 3 weeks ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆19May 6, 2024Updated 2 years ago
- Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style☆15Aug 18, 2025Updated last year
- [EMNLP 2024] IFCap: Image-like Retrieval and Frequency-based Entity Filtering for Zero-shot Captioning☆16May 13, 2025Updated last year
- A desktop UI crafted for Hermes Agent, built on Tauri 2 and React. Positioning: A native companion application for macOS, featuring zero-…☆20Jun 17, 2026Updated 3 months ago
- ☆21Jun 4, 2026Updated 4 months ago
- ☆68Apr 3, 2023Updated 3 years ago
- ☆17Jul 6, 2023Updated 3 years ago