Fine-tuning Wav2Vec2.0 on Common Voice(zh-HK)
☆16May 8, 2022Updated 4 years ago
Alternatives and similar repositories for CantoASR
Users that are interested in CantoASR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ASR project with pytorch-lightning☆20Mar 21, 2025Updated last year
- ASR: fine-tune wav2vec 2.0 with transformers☆21Sep 13, 2021Updated 4 years ago
- wav2vec2 asr with transformers☆16Oct 26, 2021Updated 4 years ago
- ☆12Aug 9, 2021Updated 5 years ago
- cantonese-mandarin unsupervised neural translation for sw project☆29May 2, 2023Updated 3 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆12Feb 9, 2021Updated 5 years ago
- BERT Tokenizer with vocabulary tailored for Cantonese☆23Oct 27, 2022Updated 3 years ago
- An audio and transcribed corpus of contemporary Hong Kong Cantonese☆41Aug 21, 2026Updated last week
- ☆16Aug 1, 2025Updated last year
- Implementation of the paper "Confidence estimation for attention based sequence to sequence models for speech recognition"☆16May 9, 2021Updated 5 years ago
- 语音识别 论文 前沿☆53Jan 8, 2022Updated 4 years ago
- PyTorch implementation of Sequence Transduction with Recurrent Neural Networks (RNN-T) speech recognition paper☆16Mar 4, 2022Updated 4 years ago
- Python scripts and datasets of the "Extremely Low-Resource Neural Machine Translation: A Case Study of Cantonese" project☆16Oct 28, 2022Updated 3 years ago
- This project focuses on the classification of animal sounds using deep learning. The core idea is to utilize audio processing techniques …☆10Dec 3, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- The Cantonese Wordnet☆15Dec 4, 2023Updated 2 years ago
- Conformer RNN-Transducer☆14May 25, 2022Updated 4 years ago
- ☆17Nov 30, 2021Updated 4 years ago
- View and compare Bible Translations in an innovative interlinear format. Run on Windows or Web.☆15Updated this week
- FunAudioLLM homepage☆17Dec 11, 2024Updated last year
- Road crack detection project based on NestedUnet model☆22Jan 22, 2022Updated 4 years ago
- Implements of CTC, Speech-Transformer and CIF for end-to-end speech recognition with pytorch☆23Jul 28, 2020Updated 6 years ago
- An English-to-Cantonese machine translation model☆55Mar 26, 2025Updated last year
- Image-to-Image Translation in PyTorch☆13Mar 2, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆41May 15, 2023Updated 3 years ago
- 基于GMM的0-9孤立词语音识别系统☆10Sep 29, 2020Updated 5 years ago
- ☆20Jun 16, 2026Updated 2 months ago
- A simple 3d sound head related transfer function (HRTF) implementation.☆23Dec 25, 2015Updated 10 years ago
- ☆23Aug 12, 2025Updated last year
- ☆15Aug 30, 2022Updated 4 years ago
- A merged version of multiple open-source German speech datasets.☆34May 3, 2024Updated 2 years ago
- A simple implementation of the paper https://arxiv.org/pdf/1910.00716v1.pdf☆31Feb 10, 2022Updated 4 years ago
- Add n-gram and large language model (LLM) support to Whisper models.☆43May 6, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- PyTorch Implementations for End-to-End Automatic Speech Recognition☆127Jun 10, 2019Updated 7 years ago
- Learning BPE embeddings by first learning a segmentation model and then training word2vec☆19Dec 18, 2022Updated 3 years ago
- End-to-end MOdeling of ASR (Automatic Speech Recognition)☆33Feb 16, 2023Updated 3 years ago
- ConfyUIで複数のLoRAを気楽に使うためのカスタムノード☆11Mar 27, 2025Updated last year
- This is a simple script to generate random musical chord progressions that make sense.☆12Mar 22, 2019Updated 7 years ago
- (R&D) Text to speech using phonemes as inputs and audio codec codes as outputs. Loosely based on MegaByte, VALL-E and Encodec.☆48Sep 4, 2023Updated 2 years ago
- ☆16Nov 11, 2025Updated 9 months ago