A repository used to organize content related to Large Speech(Audio) Model, including paper, data, applications, tools and so on.
☆28Nov 8, 2025Updated 9 months ago
Alternatives and similar repositories for Awesome-Large-Speech-Model
Users that are interested in Awesome-Large-Speech-Model are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- S3PRL for Speech Emotion Recognition (see s3prl > downstream)☆15Feb 28, 2026Updated 5 months ago
- This repo contains some extensions of deepspeed-chat for fine-tuning LLMs (SFT+RLHF).☆21Jul 2, 2024Updated 2 years ago
- Redundancy Undermines the Trustworthiness of Self-Interpretable GNNs, International Conference on Machine Learning (ICML), 2025☆15Jun 23, 2025Updated last year
- [AAAI 2026 & ACL 2026] The official implementation of the DIFFA series for dLLM-based large audio language model☆83Apr 7, 2026Updated 4 months ago
- The project for speech translation☆12Sep 28, 2023Updated 2 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- CoNeTTE: An efficient Audio Captioning system leveraging multiple datasets with Task Embedding☆23Dec 17, 2025Updated 7 months ago
- [ICASSP 2024] KNN-CTC: Enhancing ASR via Retrieval of CTC Pseudo Labels☆42Mar 20, 2024Updated 2 years ago
- SpeechBrain中文文档☆12Mar 20, 2021Updated 5 years ago
- Repo for the FB AI Speech team.☆27Aug 24, 2021Updated 4 years ago
- ☆11Nov 16, 2024Updated last year
- Paper List☆18Jul 2, 2025Updated last year
- Open-Source Turn-Taking Detection Model and Dataset for Full-Duplex Spoken Dialogue Systems☆129Jan 25, 2026Updated 6 months ago
- Jupyter notebooks for the book "Deep Learning with Python"☆11Aug 24, 2020Updated 5 years ago
- [NAACL 2024] Better Zero-Shot Reasoning with Role-Play Prompting☆36Nov 14, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official Implementation of "Prefix tuning for Automated Audio Captioning(ICASSP 2023)"☆30Dec 6, 2023Updated 2 years ago
- A torch implementation of a recursion which turns out to be useful for RNN-T.☆148Aug 25, 2023Updated 2 years ago
- 세종말뭉치 가공데이터 Repository☆14Sep 11, 2018Updated 7 years ago
- ☆15Nov 26, 2024Updated last year
- A demo to show how to convert a TensorFlow model to TensorRT uff or PLAN☆11Jul 22, 2018Updated 8 years ago
- Pitch estimation network (PiENet) for noise-robust neural F0 estimation of speech signals☆50Jul 24, 2019Updated 7 years ago
- speech-dereverberation-using-GANs☆13Jan 28, 2019Updated 7 years ago
- mnn asr demo.☆27Mar 24, 2025Updated last year
- 以Word2Vec和LSTM为基础,实现一个语言模型☆11Nov 7, 2017Updated 8 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [DATE 2023] Pipe-BD: Pipelined Parallel Blockwise Distillation☆12Jul 13, 2023Updated 3 years ago
- ☆30Apr 22, 2024Updated 2 years ago
- Zero-shot Domain-sensitive Speech Recognition with Prompt-conditioning Fine-tuning (ASRU2023)☆26Oct 10, 2023Updated 2 years ago
- Code for SLT 2016 paper on Grapheme-to-Phoneme conversion using attention based encoder-decoder models☆15Feb 20, 2019Updated 7 years ago
- TensorRT-5 based inference engine in Python☆14Sep 23, 2018Updated 7 years ago
- ☆13Apr 16, 2018Updated 8 years ago
- ☆21Mar 6, 2026Updated 5 months ago
- ☆11Aug 10, 2022Updated 4 years ago
- 记录一些常用算法的实现(涵盖常用的数据结构,机器学习以及语音识别中常用算法)☆14Jul 10, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- NAR-BERT-ASR☆10Sep 27, 2021Updated 4 years ago
- The MIR-MLPop dataset and the official implementation of the paper "MIR-MLPop: A Multilingual Pop Music Dataset with Time-Aligned Lyrics …☆35Apr 22, 2024Updated 2 years ago
- ☆16Nov 5, 2018Updated 7 years ago
- 命名实体识别(NER),分词(CWS),实体分类(Entity Typing),关系抽取(Relation Extraction)等任务数据集整理☆13Mar 7, 2020Updated 6 years ago
- Official implementation of the paper titled "Age and Gender Recognition Using a Convolutional Neural Network with a Specially Designed Mu…☆28Mar 5, 2024Updated 2 years ago
- (已过时)WaveNet 声码器☆21Mar 5, 2020Updated 6 years ago
- ☆15Sep 13, 2022Updated 3 years ago