基于python的语音识别服务部署,任何一个支持一句话解码的ASR模型接口,都可仿照该框架部署自己的语音识别服务
☆55Mar 8, 2022Updated 4 years ago
Alternatives and similar repositories for ASR_python_deploy
Users that are interested in ASR_python_deploy are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 实现对语音进行端点检测,并去除语音中静音段,可以作为语音信号处理的一个预处理☆18Jul 19, 2021Updated 5 years ago
- Generative_Annotation_NEC: A novel NEC method that utilizes speech sound features to retrieve candidate entities and a generative method …☆17Dec 2, 2025Updated 8 months ago
- 任务型对话系统(Task-based Dialogue System)☆66Feb 8, 2022Updated 4 years ago
- A Framework for Multimodal Parsing, Contextual Narration, and Hierarchical Labeling of ESG Reports☆17Nov 14, 2025Updated 9 months ago
- using microphone☆16Sep 2, 2021Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- The project is associated with the recently-launched ICASSP 2022 Multi-channel Multi-party Meeting Transcription Challenge (M2MeT) to pro…☆143Jun 10, 2022Updated 4 years ago
- ☆19Jan 5, 2020Updated 6 years ago
- ☆20Nov 22, 2020Updated 5 years ago
- Keyword spotting for audio with attention (KWS model for audio)☆18Jul 15, 2021Updated 5 years ago
- Recipe for LibriPhrase☆38Sep 2, 2023Updated 2 years ago
- A ctc decoder for both online and offline asr model☆66Nov 18, 2023Updated 2 years ago
- Framework for Detection Evaluation (F4DE) : set of evaluation tools for detection evaluations and for specific NIST-coordinated evaluatio…☆26Jul 6, 2017Updated 9 years ago
- Implementation of the paper: Channel-wise Gated Res2Net: Towards Robust Detection of Synthetic Speech Attacks (INTERSPEECH 2021)☆32Jul 21, 2021Updated 5 years ago
- [ICASSP'22] Integer-only Zero-shot Quantization for Efficient Speech Recognition☆34Oct 11, 2021Updated 4 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- PyTorch implementation of LiMuSE☆33Oct 11, 2022Updated 3 years ago
- ☆33Aug 6, 2021Updated 5 years ago
- EfficientNet-Absolute Zero for Continuous Speech Keyword Spotting☆23Jun 16, 2022Updated 4 years ago
- Production First and Production Ready End-to-End Speech Recognition Toolkit☆5,224Jun 15, 2026Updated 2 months ago
- PyTorch implementation of TinyWASE described in our paper "Compressing Speaker Extraction Model with Ultra-low Precision Quantization and…☆11Jun 28, 2021Updated 5 years ago
- PolEval 2021 Task 1☆15Jun 28, 2022Updated 4 years ago
- Chinese Prosodic Structure Prediction☆10May 18, 2019Updated 7 years ago
- HyperTMO: A trusted multi-omics integration framework based on Hypergraph convolutional network for patient classification☆13Apr 1, 2024Updated 2 years ago
- Performed document clustering using the DBSCAN clustering algorithm☆14Oct 21, 2020Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆171Nov 28, 2024Updated last year
- This repository is for the paper Incorporating External POS Tagger for Punctuation Restoration. Proc. Interspeech 2021, 1987-1991, doi: 1…☆11May 24, 2026Updated 2 months ago
- A library for adding punctuation into a text from ASR.☆20May 8, 2023Updated 3 years ago
- Dual-level deep evidential fusion☆12Mar 19, 2024Updated 2 years ago
- Simple Kaldi model server for chain (nnet3) models in online recognition mode directly from a local microphone☆35Feb 18, 2022Updated 4 years ago
- simple energy vad☆19Jun 3, 2017Updated 9 years ago
- CosyVoice_DPO_NOTES: Supercharge Your Cosyvoice model with Cutting-Edge DPO Fine-Tuning!☆127Aug 8, 2025Updated last year
- 算法导论☆10Dec 20, 2021Updated 4 years ago
- Code for Interspeech2022 paper DeID-VC: Speaker De-identification via Zero-shot Pseudo Voice Conversion☆13May 6, 2023Updated 3 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Monotonic Alignment Search☆100Jun 9, 2025Updated last year
- Official repository for "Structure-Enhanced Pop Music Generation via Harmony-Aware Learning", ACM MM 2022.☆14Mar 22, 2023Updated 3 years ago
- ☆11Oct 31, 2020Updated 5 years ago
- Rainbow Keywords - Official PyTorch Implementation☆14Jun 27, 2024Updated 2 years ago
- ☆10Dec 11, 2021Updated 4 years ago
- 百度汉语爬虫,爬取unicode字符集中所有汉字及所有汉字所成所有词的信息,信息包括拼音、释义、百科释义、英文翻译☆11Dec 23, 2019Updated 6 years ago
- MMM 2021: Crossed-Time Delay Neural Network for Speaker Recognition☆11Dec 4, 2021Updated 4 years ago