Fine-tuning Wav2Vec2.0 on Common Voice(zh-HK)
☆16May 8, 2022Updated 4 years ago
Alternatives and similar repositories for CantoASR
Users that are interested in CantoASR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- wav2vec2 asr with transformers☆16Oct 26, 2021Updated 4 years ago
- ASR, End-to-End, end2end, Speech Recognition, 端到端语音识别☆12Oct 25, 2020Updated 5 years ago
- ☆12Aug 9, 2021Updated 4 years ago
- cantonese-mandarin unsupervised neural translation for sw project☆29May 2, 2023Updated 3 years ago
- BERT Tokenizer with vocabulary tailored for Cantonese☆23Oct 27, 2022Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- An audio and transcribed corpus of contemporary Hong Kong Cantonese☆41Dec 30, 2020Updated 5 years ago
- ☆16Aug 1, 2025Updated 11 months ago
- Implementation of the paper "Confidence estimation for attention based sequence to sequence models for speech recognition"☆16May 9, 2021Updated 5 years ago
- 语音识别 论文 前沿☆53Jan 8, 2022Updated 4 years ago
- Final training script from HuggingFace Whisper Fine tuning event - to get best results on finetuned model.☆12Dec 24, 2022Updated 3 years ago
- Python scripts and datasets of the "Extremely Low-Resource Neural Machine Translation: A Case Study of Cantonese" project☆16Oct 28, 2022Updated 3 years ago
- This project focuses on the classification of animal sounds using deep learning. The core idea is to utilize audio processing techniques …☆10Dec 3, 2024Updated last year
- The Cantonese Wordnet☆15Dec 4, 2023Updated 2 years ago
- Conformer RNN-Transducer☆14May 25, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆17Nov 30, 2021Updated 4 years ago
- View and compare Bible Translations in an innovative interlinear format. Run on Windows or Web.☆15Updated this week
- FunAudioLLM homepage☆17Dec 11, 2024Updated last year
- Implementaion RNN tranceducer☆23Jun 25, 2019Updated 7 years ago
- Implements of CTC, Speech-Transformer and CIF for end-to-end speech recognition with pytorch☆23Jul 28, 2020Updated 5 years ago
- An English-to-Cantonese machine translation model☆55Mar 26, 2025Updated last year
- Image-to-Image Translation in PyTorch☆13Mar 2, 2021Updated 5 years ago
- ☆41May 15, 2023Updated 3 years ago
- 基于GMM的0-9孤立词语音识别系统☆10Sep 29, 2020Updated 5 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- This repository is a Python implementation of HMM-DNN model.☆15Jul 3, 2020Updated 6 years ago
- ☆17Jun 16, 2026Updated last month
- A simple 3d sound head related transfer function (HRTF) implementation.☆23Dec 25, 2015Updated 10 years ago
- ☆22Aug 12, 2025Updated 11 months ago
- ☆15Aug 30, 2022Updated 3 years ago
- A merged version of multiple open-source German speech datasets.☆34May 3, 2024Updated 2 years ago
- A simple implementation of the paper https://arxiv.org/pdf/1910.00716v1.pdf☆31Feb 10, 2022Updated 4 years ago
- Add n-gram and large language model (LLM) support to Whisper models.☆43May 6, 2025Updated last year
- This app is intended to automatically create a corpus for ASR systems using pseudo-labeling.☆27Feb 15, 2024Updated 2 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- PyTorch Implementations for End-to-End Automatic Speech Recognition☆127Jun 10, 2019Updated 7 years ago
- Learning BPE embeddings by first learning a segmentation model and then training word2vec☆19Dec 18, 2022Updated 3 years ago
- End-to-end MOdeling of ASR (Automatic Speech Recognition)☆33Feb 16, 2023Updated 3 years ago
- ☆18May 28, 2024Updated 2 years ago
- (R&D) Text to speech using phonemes as inputs and audio codec codes as outputs. Loosely based on MegaByte, VALL-E and Encodec.☆48Sep 4, 2023Updated 2 years ago
- rime-cantonese 上游詞表倉庫☆33Jun 29, 2026Updated 3 weeks ago
- ☆16Nov 11, 2025Updated 8 months ago