Montreal Forced Aligner for Vietnamese
☆15Oct 23, 2023Updated 2 years ago
Alternatives and similar repositories for ViMFA
Users that are interested in ViMFA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- LIGHTVOC AN UPSAMPLING-FREE GAN VOCODER BASED ON CONFORMER AND INVERSE SHORT-TIME FOURIER TRANSFORM☆18May 17, 2024Updated 2 years ago
- Vi_G2P or ViG2P: G2P package for Vietnamese: based on vPhon and phonology knowledge to convert Raw text - Graphoneme to IPA☆108Jun 21, 2024Updated 2 years ago
- Python - NSW package for Vietnamese: Normalization system to convert numbers, abbreviations, and words that cannot be pronounced into syl…☆67Jan 1, 2025Updated last year
- Simple text to phonemes converter for multiple languages☆20Nov 21, 2022Updated 3 years ago
- ☆15Jul 24, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Generate degraded speech datasets for noise-robust ASR benchmarking☆45Jun 9, 2026Updated 2 months ago
- An official implementation of Style-Talker for Spoken Dialogue Generation☆23Jan 12, 2025Updated last year
- A Vietnamese phonetizer☆56May 29, 2024Updated 2 years ago
- Vietnamese Voice Cloning System using Speaker Verification training on multispeaker VITS☆56Dec 1, 2023Updated 2 years ago
- StyleTTS-ZS: Efficient High-Quality Zero-Shot Text-to-Speech Synthesis with Distilled Time-Varying Style Diffusion☆10Sep 22, 2024Updated last year
- Sing any popular song with your voice☆11Jul 10, 2022Updated 4 years ago
- Dictionary of pairs of Korean word and IPA crawled from Wiktionary (Korean edition)☆23Nov 12, 2025Updated 9 months ago
- Once more Diarization: Improving meeting transcription systems through segment-level speaker reassignment☆14Feb 5, 2025Updated last year
- A repo containing download guidance and corresponding scripts of the VoxBlink dataset.☆30Apr 16, 2024Updated 2 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- ☆38Oct 10, 2025Updated 10 months ago
- End-to-End Chinese Speaker Identification☆11Nov 17, 2022Updated 3 years ago
- Text frontend for ESPnet tts recipes☆35Jun 1, 2021Updated 5 years ago
- This repository contains source codes for SoftCTC. Original paper can be found here: https://arxiv.org/abs/2212.02135☆19Mar 7, 2023Updated 3 years ago
- A Vietnamese Text-to-Speech library that provides high-quality speech synthesis with voice cloning capabilities☆103Jul 14, 2025Updated last year
- [ACL 2025] OZSpeech: One-step Zero-shot Speech Synthesis with Learned-Prior-Conditioned Flow Matching☆45Feb 9, 2025Updated last year
- Code for "Weakly-supervised Fingerspelling Recognition in British Sign Language Videos", BMVC 2022.☆12Jun 22, 2023Updated 3 years ago
- ☆55Sep 3, 2025Updated 11 months ago
- A synthesized dataset for Vietnamese TTS task☆66May 6, 2022Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆13Apr 12, 2025Updated last year
- Dự án công cụ chuyển đổi giọng nói dành cho người Việt☆33Aug 5, 2026Updated 3 weeks ago
- [ICASSP'23] Online speaker clustering☆19Feb 22, 2026Updated 6 months ago
- ViSen is library to format tone of Vietnamese sentences☆22Nov 9, 2021Updated 4 years ago
- sherpa with mlx☆15Aug 2, 2025Updated last year
- ☆10Nov 27, 2024Updated last year
- ViStreamASR - Real-Time Vietnamese Speech Recognition☆61Jul 12, 2025Updated last year
- ☆19May 6, 2024Updated 2 years ago
- Handling audio files in Python☆39Aug 5, 2026Updated 3 weeks ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Model Detect Phishing Website by using Machine Learning for Project Chongluadao☆18Aug 10, 2021Updated 5 years ago
- Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style☆15Aug 18, 2025Updated last year
- [EMNLP 2024] IFCap: Image-like Retrieval and Frequency-based Entity Filtering for Zero-shot Captioning☆16May 13, 2025Updated last year
- A lightweight, efficient variation of the StyleTTS 2 text‐to‐speech model.☆51May 22, 2025Updated last year
- A desktop UI crafted for Hermes Agent, built on Tauri 2 and React. Positioning: A native companion application for macOS, featuring zero-…☆21Jun 17, 2026Updated 2 months ago
- ☆21Jun 4, 2026Updated 2 months ago
- ☆68Apr 3, 2023Updated 3 years ago