CTC decoder with hotwords for ASR.
☆38Jun 15, 2026Updated last month
Alternatives and similar repositories for asr-decoder
Users that are interested in asr-decoder are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- One command to start a streaming ASR server.☆12Oct 2, 2024Updated last year
- An N-gram punctuator for Chinese and English.☆18Oct 14, 2025Updated 9 months ago
- ☆23Oct 17, 2024Updated last year
- Offline Speaker Diarization with SenseVoice by Sherpa ONNX.☆15Dec 23, 2024Updated last year
- Pseudo Streaming SenseVoice with Hotwords☆466Jun 15, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- One command to build TLG.fst for WeNet.☆30Oct 11, 2022Updated 3 years ago
- faster inference☆27Jan 20, 2025Updated last year
- paraformer(chinense asr) online onnx runtime for python☆54Mar 27, 2024Updated 2 years ago
- a Neural Vocoder supporting Ring Attention, Conformer and NSF.☆25Aug 1, 2025Updated 11 months ago
- (WIP)long form speech generatoins☆30Apr 2, 2025Updated last year
- Compute WER and SER for speech recognition evaluation☆27Jun 6, 2026Updated last month
- ☆33Feb 4, 2025Updated last year
- Train no-reference speech quality estimators with multiple datasets via learned, per-dataset alignments.☆18Aug 1, 2025Updated 11 months ago
- Torch Audio Forced Aligner for Mixed Chinese (Mandarin or Cantonese) and English.☆61Sep 5, 2025Updated 10 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Forced alignment decoder for Whisper.☆16Mar 13, 2024Updated 2 years ago
- low-latency realtime ASR based on FireRedASR☆62Jul 8, 2025Updated last year
- sherpa with mlx☆15Aug 2, 2025Updated 11 months ago
- ☆35Sep 6, 2025Updated 10 months ago
- Speech-To-Text forced-alignment Speech processing Universal PERformance Benchmark☆39May 7, 2025Updated last year
- Streaming Vocos☆31Jun 10, 2025Updated last year
- Code associated with the paper: CTC-DRO: Robust Optimization for Reducing Language Disparities in Speech Recognition.☆17May 16, 2025Updated last year
- This repository contains prompts & best practices to annotate audio clips with a very high degree of details using Audio-Language-Models☆35Oct 13, 2024Updated last year
- A playground for experimenting with acoustic echo cancellation using a microphone, speaker, and ONNX.☆13Oct 22, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Ultra-low-bitrate Speech Codec for Speech Language Modeling Applications☆92Dec 20, 2024Updated last year
- Simple voice activity detection (VAD) algorithm in Python☆15Aug 10, 2023Updated 2 years ago
- ☆16Nov 9, 2023Updated 2 years ago
- noise reduction☆17Jul 3, 2024Updated 2 years ago
- 来自于文章Paraformer-v2: An improved non-autoregressive transformer for noise-robust speech recognition☆29Nov 20, 2024Updated last year
- ☆19Mar 22, 2024Updated 2 years ago
- A simple command line tool to calculate WER for ASR.☆14Oct 14, 2024Updated last year
- Simplistic Implementation of Zipformer:A faster and better encoder for automatic speech recognition in PyTorch☆22Jun 3, 2024Updated 2 years ago
- A unified tokenizer that is capable of both extracting semantic information and enabling high-fidelity audio reconstruction.☆145Sep 19, 2025Updated 10 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Meanflow and multilingual for F5-TTS model☆16Aug 23, 2025Updated 11 months ago
- Official PyTorch inference code for the Interspeech 2025 paper: Efficient Speech Enhancement via Embeddings from Pre-trained Generative A…☆81Jun 16, 2025Updated last year
- Text-to-text alignment algorithm for speech recognition error analysis.☆32Jun 23, 2026Updated last month
- (Interspeech 2023 & ICASSP 2024) Official repository for ARMHuBERT and STaRHuBERT☆41Aug 29, 2024Updated last year
- This is a repository for fine-tuning Qwen2-Audio, currently supporting Distributed Data Parallel (DDP) and DeepSpeed.☆50Jul 28, 2025Updated 11 months ago
- semantic tokenizer for speech and music☆20Jul 6, 2025Updated last year
- mnn asr demo.☆27Mar 24, 2025Updated last year