Montreal Forced Aligner for Vietnamese
☆15Oct 23, 2023Updated 2 years ago
Alternatives and similar repositories for ViMFA
Users that are interested in ViMFA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- LIGHTVOC AN UPSAMPLING-FREE GAN VOCODER BASED ON CONFORMER AND INVERSE SHORT-TIME FOURIER TRANSFORM☆18May 17, 2024Updated 2 years ago
- Python - NSW package for Vietnamese: Normalization system to convert numbers, abbreviations, and words that cannot be pronounced into syl…☆67Jan 1, 2025Updated last year
- Simple text to phonemes converter for multiple languages☆20Nov 21, 2022Updated 3 years ago
- ☆15Jul 24, 2025Updated last year
- Generate degraded speech datasets for noise-robust ASR benchmarking☆45Jun 9, 2026Updated 3 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- An official implementation of Style-Talker for Spoken Dialogue Generation☆23Jan 12, 2025Updated last year
- Vietnamese Voice Cloning System using Speaker Verification training on multispeaker VITS☆55Dec 1, 2023Updated 2 years ago
- A Vietnamese phonetizer☆56May 29, 2024Updated 2 years ago
- StyleTTS-ZS: Efficient High-Quality Zero-Shot Text-to-Speech Synthesis with Distilled Time-Varying Style Diffusion☆10Sep 22, 2024Updated last year
- Sing any popular song with your voice☆11Jul 10, 2022Updated 4 years ago
- Dictionary of pairs of Korean word and IPA crawled from Wiktionary (Korean edition)☆24Nov 12, 2025Updated 10 months ago
- Once more Diarization: Improving meeting transcription systems through segment-level speaker reassignment☆14Feb 5, 2025Updated last year
- A repo containing download guidance and corresponding scripts of the VoxBlink dataset.☆30Apr 16, 2024Updated 2 years ago
- ☆38Oct 10, 2025Updated 11 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- End-to-End Chinese Speaker Identification☆11Nov 17, 2022Updated 3 years ago
- Text frontend for ESPnet tts recipes☆35Sep 2, 2026Updated 2 weeks ago
- This repository contains source codes for SoftCTC. Original paper can be found here: https://arxiv.org/abs/2212.02135☆19Mar 7, 2023Updated 3 years ago
- A Vietnamese Text-to-Speech library that provides high-quality speech synthesis with voice cloning capabilities☆102Jul 14, 2025Updated last year
- [ACL 2025] OZSpeech: One-step Zero-shot Speech Synthesis with Learned-Prior-Conditioned Flow Matching☆45Feb 9, 2025Updated last year
- Code for "Weakly-supervised Fingerspelling Recognition in British Sign Language Videos", BMVC 2022.☆12Jun 22, 2023Updated 3 years ago
- ☆55Sep 3, 2025Updated last year
- A synthesized dataset for Vietnamese TTS task☆66May 6, 2022Updated 4 years ago
- ☆13Apr 12, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Dự án công cụ chuyển đổi giọng nói dành cho người Việt☆33Sep 1, 2026Updated 2 weeks ago
- [ICASSP'23] Online speaker clustering☆19Feb 22, 2026Updated 6 months ago
- ViSen is library to format tone of Vietnamese sentences☆22Nov 9, 2021Updated 4 years ago
- sherpa with mlx☆15Aug 2, 2025Updated last year
- ☆10Nov 27, 2024Updated last year
- ViStreamASR - Real-Time Vietnamese Speech Recognition☆61Jul 12, 2025Updated last year
- Handling audio files in Python☆39Updated this week
- ☆19May 6, 2024Updated 2 years ago
- Model Detect Phishing Website by using Machine Learning for Project Chongluadao☆18Aug 10, 2021Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style☆15Aug 18, 2025Updated last year
- [EMNLP 2024] IFCap: Image-like Retrieval and Frequency-based Entity Filtering for Zero-shot Captioning☆16May 13, 2025Updated last year
- A lightweight, efficient variation of the StyleTTS 2 text‐to‐speech model.☆53May 22, 2025Updated last year
- A desktop UI crafted for Hermes Agent, built on Tauri 2 and React. Positioning: A native companion application for macOS, featuring zero-…☆21Jun 17, 2026Updated 3 months ago
- ☆21Jun 4, 2026Updated 3 months ago
- ☆68Apr 3, 2023Updated 3 years ago
- ☆17Jul 6, 2023Updated 3 years ago