Transformation spoken text to written text
☆31May 14, 2024Updated 2 years ago
Alternatives and similar repositories for spoken-norm
Users that are interested in spoken-norm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Vietnamese self-supervised Wav2vec2 model☆61Nov 5, 2022Updated 3 years ago
- Cải thiện Elasticsearch trong bài toán semantic search sử dụng phương pháp Sentence Embeddings☆25May 27, 2021Updated 5 years ago
- ICONIP2021 - A Vietnamese Medical Dataset for IC and NER☆24Aug 8, 2023Updated 2 years ago
- Vietnamese song lyric alignment framework☆68Dec 11, 2022Updated 3 years ago
- A Robustly Optimized BERT Pretraining Approach for Vietnamese☆32Jul 25, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- End-to-End Vietnamese Speech Recognition using wav2vec 2.0☆106Sep 3, 2021Updated 4 years ago
- Xây dựng tập dữ liệu 500GB (20% done) văn bản tiếng Việt để huấn luyện mô hình ngôn ngữ lớn☆29Apr 7, 2023Updated 3 years ago
- Pre-training script for BART in JAX/Flax☆37Aug 4, 2022Updated 3 years ago
- Machine Reading Comprehension special for the Vietnamese language☆41Mar 13, 2022Updated 4 years ago
- Implementation of paper: ConvNet for the 2020s☆21Oct 13, 2022Updated 3 years ago
- Electra pre-trained model using Vietnamese corpus☆65Jun 12, 2023Updated 3 years ago
- ViSen is library to format tone of Vietnamese sentences☆22Nov 9, 2021Updated 4 years ago
- Submission for AIviVN sentiment analysis contest https://www.aivivn.com/contests/1☆15Oct 12, 2021Updated 4 years ago
- Build English-Vietnamese machine translation with ProtonX Transformer. :D☆75Sep 13, 2021Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- (WIP) Parallel inference for black-forest-labs' FLUX model.☆19Nov 18, 2024Updated last year
- Wenet speech to text for react native☆10Nov 1, 2022Updated 3 years ago
- Mã nguồn của mô hình xếp hạng 7 tại cuộc thi phân loại sắc thái bình luận do Aivivn tổ chức.☆12Mar 21, 2019Updated 7 years ago
- Vietnamese Punctuation Prediction using Pretrained Language Models☆14May 8, 2022Updated 4 years ago
- paraphase sentence☆11Aug 22, 2025Updated 11 months ago
- ☆13Oct 27, 2021Updated 4 years ago
- acnn for text-independent speaker recognition☆10Feb 8, 2022Updated 4 years ago
- 44100Hz日本語音源に対応させた unofficial vits2-TTS implementation in pytorchです。☆24Sep 1, 2023Updated 2 years ago
- ☆10Sep 19, 2018Updated 7 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆20Mar 25, 2025Updated last year
- A SPMI Lab toolkit for language models.☆11Apr 12, 2017Updated 9 years ago
- Repository to track the progress in Vietnamese Natural Language Processing, including the datasets and the current state-of-the-art for t…☆376Sep 5, 2022Updated 3 years ago
- Automatic Speech Recognition (ASR) system for the Samrómur speech corpus using Kaldi☆12Sep 30, 2022Updated 3 years ago
- [IJCAI'23] Learning to Speak from Text for Low-Resource TTS☆65May 30, 2023Updated 3 years ago
- NMT based punctuation prediction system using lexical and acoustic features .☆14Mar 30, 2020Updated 6 years ago
- Speaker Diarization library in Python. Performs VAD, Segmentation, Linear Clustering, Hierarchical Clustering☆15Jul 28, 2017Updated 9 years ago
- Ai cũng có thể tự tạo chatbot bằng huấn luyện chỉ dẫn, với 12G GPU (RTX 3060) và khoảng vài chục MB dữ liệu☆112Jun 10, 2023Updated 3 years ago
- Huggingface Implementation of AV-HuBERT on the MuAViC Dataset☆19Mar 6, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- S3PRL for Speech Emotion Recognition (see s3prl > downstream)☆15Feb 28, 2026Updated 5 months ago
- NISQA - Non-Intrusive Speech Quality and TTS Naturalness Assessment☆16Apr 13, 2022Updated 4 years ago
- 将normalize过的中文文本,做逆向normalize。具体功能即实现 chinese_text_normalization的逆向版本。☆13Apr 7, 2021Updated 5 years ago
- Fast whitespace correction with Transformers☆18Aug 22, 2025Updated 11 months ago
- Forced alignment decoder for Whisper.☆16Mar 13, 2024Updated 2 years ago
- Speech-To-Text forced-alignment Speech processing Universal PERformance Benchmark☆39May 7, 2025Updated last year
- ☆35Aug 27, 2021Updated 4 years ago