A curated list of awesome papers on contextualizing E2E ASR outputs
☆81May 10, 2023Updated 3 years ago
Alternatives and similar repositories for awesome-asr-contextualization
Users that are interested in awesome-asr-contextualization are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- open-source Mandarian biased word dataset☆14Sep 21, 2023Updated 2 years ago
- Zero-shot Domain-sensitive Speech Recognition with Prompt-conditioning Fine-tuning (ASRU2023)☆26Oct 10, 2023Updated 2 years ago
- ☆88Jul 31, 2025Updated 11 months ago
- Repo for the FB AI Speech team.☆27Aug 24, 2021Updated 4 years ago
- ☆17May 5, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICASSP 2022] Improving End-to-End Contextual Speech Recognition with Fine-Grained Contextual Knowledge Selection☆25Jul 14, 2026Updated 2 weeks ago
- Memory efficient transducer loss computation☆70Jun 10, 2022Updated 4 years ago
- ☆20Jun 3, 2024Updated 2 years ago
- Decoders from Kaldi using OpenFst☆35Apr 10, 2026Updated 3 months ago
- A simple command line tool to calculate WER for ASR.☆14Updated this week
- [ICASSP2023] Source code, model links and open test sets for paper SeACo-Paraformer.☆44Mar 15, 2024Updated 2 years ago
- Dataset for Pinyin Regularization in Error Correction for Chinese Speech Recognition with Large Language Models in Interspeech 2024.☆16Jul 4, 2024Updated 2 years ago
- ☆37Jun 9, 2026Updated last month
- ☆18Jul 22, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [INTERSPEECH 2023] Knowledge Transfer from Pre-trained Language Models to Cif-based Recognizers via Hierarchical Distillation☆41Jul 14, 2026Updated 2 weeks ago
- ☆11Nov 5, 2021Updated 4 years ago
- E2E system with LF-MMI; word N-gram for Mandarin☆167Apr 29, 2022Updated 4 years ago
- Some fast-ish algorithms for batch text search in moderate-sized collections, intended for data cleanup☆79Jun 30, 2025Updated last year
- Properly handle position-dependent phones in a subword lexicon FST☆31Oct 26, 2020Updated 5 years ago
- ☆14Jun 17, 2024Updated 2 years ago
- Small compression utility☆38Jan 20, 2026Updated 6 months ago
- This Repository surveys the paper focusing on Prompting and Adapters for Speech Processing.☆113Aug 4, 2023Updated 2 years ago
- Automatic Speech Recognition (ASR) system for the Samrómur speech corpus using Kaldi☆12Sep 30, 2022Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [EMNLP 2025 Findings] A complete cross-modal RAG system for end-to-end speech-to-speech large models, including ASR-based Retrieval and E…☆31Jul 11, 2025Updated last year
- End-to-end ASR/LM implementation with PyTorch☆594Aug 30, 2021Updated 4 years ago
- Python wrapper for kaldi's arpa2fst☆38Aug 27, 2025Updated 11 months ago
- PolEval 2021 Task 1☆15Jun 28, 2022Updated 4 years ago
- End-to-end MOdeling of ASR (Automatic Speech Recognition)☆33Feb 16, 2023Updated 3 years ago
- ☆13Mar 30, 2023Updated 3 years ago
- ☆28Apr 24, 2026Updated 3 months ago
- Framework for Detection Evaluation (F4DE) : set of evaluation tools for detection evaluations and for specific NIST-coordinated evaluatio…☆26Jul 6, 2017Updated 9 years ago
- Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style☆15Aug 18, 2025Updated 11 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- sherpa with mlx☆15Aug 2, 2025Updated 11 months ago
- ☆11Oct 24, 2022Updated 3 years ago
- 语音识别 论文 前沿☆53Jan 8, 2022Updated 4 years ago
- ☆15Aug 25, 2022Updated 3 years ago
- A toolkit for Spoken Language Understanding Evaluation (SLUE) benchmark. Refer paper https://arxiv.org/abs/2111.10367 for more details. O…☆65Feb 26, 2024Updated 2 years ago
- ☆116Jun 29, 2026Updated last month
- PyTorch implementation of LF-MMI for End-to-end ASR☆221Jan 14, 2021Updated 5 years ago