The dataset of Speech Recognition
☆463Jan 4, 2026Updated 6 months ago
Alternatives and similar repositories for speech_dataset
Users that are interested in speech_dataset are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 语音方向实验室/公司/资源/实习等,欢迎推荐或自荐☆607Nov 13, 2024Updated last year
- Papers of ASR, Tools of ASR☆41Feb 14, 2025Updated last year
- Towards hot directions in industrial end to end speech recognition☆329Nov 30, 2021Updated 4 years ago
- Production First and Production Ready End-to-End Speech Recognition Toolkit☆5,172Jun 15, 2026Updated last month
- SpeechIO Leaderboard: a large, robust, comprehensive, benchmarking platform for Automatic Speech Recognition.☆547Mar 29, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- CAT is more than a CRF-based ASR toolkit: it provides a complete workflow for data-efficient end-to-end ASR, supporting CTC, CTC-CRF, RNN…☆368Feb 5, 2026Updated 5 months ago
- Paper, Code and Statistics for Self-Supervised Learning and Pre-Training on Speech.☆212Jan 18, 2024Updated 2 years ago
- Chinese text normalization for speech processing☆733Mar 18, 2023Updated 3 years ago
- An Open Source Tools for Speaker Recognition☆638Aug 5, 2024Updated last year
- Kaldi-compatible online & offline feature extraction with PyTorch, supporting CUDA, batch processing, chunk processing, and autograd - P…☆215Jul 10, 2026Updated last week
- chinese speech pretrained models☆1,209Aug 23, 2024Updated last year
- We Speech Toolkit, LLM based Speech Toolkit for Speech Understanding, Generation, and Interaction☆206Updated this week
- The repo provides information about KeSpeech Mandarin dialect dataset.☆183Oct 13, 2022Updated 3 years ago
- 3M: Multi-loss, Multi-path and Multi-level Neural Networks for speech recognition☆119Jun 22, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Large, modern dataset for speech recognition☆731Feb 26, 2024Updated 2 years ago
- The project is associated with the recently-launched ICASSP 2022 Multi-channel Multi-party Meeting Transcription Challenge (M2MeT) to pro…☆142Jun 10, 2022Updated 4 years ago
- Torch Audio Forced Aligner for Mixed Chinese (Mandarin or Cantonese) and English.☆61Sep 5, 2025Updated 10 months ago
- Production First and Production Ready End-to-End Keyword Spotting Toolkit☆740Jun 15, 2026Updated last month
- ☆1,454Updated this week
- A ctc decoder for both online and offline asr model☆66Nov 18, 2023Updated 2 years ago
- FSA/FST algorithms, differentiable, with PyTorch compatibility.☆1,348Jul 11, 2026Updated last week
- Research and Production Oriented Speaker Verification, Recognition and Diarization Toolkit☆1,359Jul 8, 2026Updated last week
- A 10000+ hours dataset for Chinese speech recognition☆621Jan 9, 2026Updated 6 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- A streaming audio reader, processor, and writer built on top of soundfile, and PyAV (bindings for FFmpeg)☆39Mar 31, 2026Updated 3 months ago
- One command to build TLG.fst for WeNet.