Montreal Forced Aligner for Vietnamese
☆15Oct 23, 2023Updated 2 years ago
Alternatives and similar repositories for ViMFA
Users that are interested in ViMFA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- LIGHTVOC AN UPSAMPLING-FREE GAN VOCODER BASED ON CONFORMER AND INVERSE SHORT-TIME FOURIER TRANSFORM☆18May 17, 2024Updated 2 years ago
- Vi_G2P or ViG2P: G2P package for Vietnamese: based on vPhon and phonology knowledge to convert Raw text - Graphoneme to IPA☆108Jun 21, 2024Updated 2 years ago
- Python - NSW package for Vietnamese: Normalization system to convert numbers, abbreviations, and words that cannot be pronounced into syl…☆67Jan 1, 2025Updated last year
- Simple text to phonemes converter for multiple languages☆20Nov 21, 2022Updated 3 years ago
- ☆15Jul 24, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Generate degraded speech datasets for noise-robust ASR benchmarking☆45Jun 9, 2026Updated 2 months ago
- An official implementation of Style-Talker for Spoken Dialogue Generation☆23Jan 12, 2025Updated last year
- A Vietnamese phonetizer☆54May 29, 2024Updated 2 years ago
- Vietnamese Voice Cloning System using Speaker Verification training on multispeaker VITS☆56Dec 1, 2023Updated 2 years ago
- StyleTTS-ZS: Efficient High-Quality Zero-Shot Text-to-Speech Synthesis with Distilled Time-Varying Style Diffusion☆10Sep 22, 2024Updated last year
- Sing any popular song with your voice☆11Jul 10, 2022Updated 4 years ago
- Dictionary of pairs of Korean word and IPA crawled from Wiktionary (Korean edition)☆23Nov 12, 2025Updated 8 months ago
- Once more Diarization: Improving meeting transcription systems through segment-level speaker reassignment☆14Feb 5, 2025Updated last year
- A repo containing download guidance and corresponding scripts of the VoxBlink dataset.☆31Apr 16, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆38Oct 10, 2025Updated 10 months ago
- End-to-End Chinese Speaker Identification☆11Nov 17, 2022Updated 3 years ago
- Text frontend for ESPnet tts recipes☆35Jun 1, 2021Updated 5 years ago
- This repository contains source codes for SoftCTC. Original paper can be found here: https://arxiv.org/abs/2212.02135☆19Mar 7, 2023Updated 3 years ago
- A Vietnamese Text-to-Speech library that provides high-quality speech synthesis with voice cloning capabilities☆105Jul 14, 2025Updated last year
- [ACL 2025] OZSpeech: One-step Zero-shot Speech Synthesis with Learned-Prior-Conditioned Flow Matching☆45Feb 9, 2025Updated last year
- Code for "Weakly-supervised Fingerspelling Recognition in British Sign Language Videos", BMVC 2022.☆12Jun 22, 2023Updated 3 years ago
- ☆53Sep 3, 2025Updated 11 months ago
- A synthesized dataset for Vietnamese TTS task☆66May 6, 2022Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆13Apr 12, 2025Updated last year
- Dự án công cụ chuyển đổi giọng nói dành cho người Việt☆33Updated this week
- [ICASSP'23] Online speaker clustering☆19Feb 22, 2026Updated 5 months ago
- ViSen is library to format tone of Vietnamese sentences☆22Nov 9, 2021Updated 4 years ago
- sherpa with mlx☆15Aug 2, 2025Updated last year
- ☆10Nov 27, 2024Updated last year
- ViStreamASR - Real-Time Vietnamese Speech Recognition☆61Jul 12, 2025Updated last year
- A library for multilingual word, phrase and sentence segmentation.☆16Jul 24, 2026Updated 2 weeks ago
- ☆19May 6, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Handling audio files in Python☆39Updated this week
- Model Detect Phishing Website by using Machine Learning for Project Chongluadao☆18Aug 10, 2021Updated 5 years ago
- Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style☆15Aug 18, 2025Updated 11 months ago
- [EMNLP 2024] IFCap: Image-like Retrieval and Frequency-based Entity Filtering for Zero-shot Captioning☆15May 13, 2025Updated last year
- A lightweight, efficient variation of the StyleTTS 2 text‐to‐speech model.☆50May 22, 2025Updated last year
- A desktop UI crafted for Hermes Agent, built on Tauri 2 and React. Positioning: A native companion application for macOS, featuring zero-…☆21Jun 17, 2026Updated last month
- ☆21Jun 4, 2026Updated 2 months ago