基于FunASR实现语音识别,包含常规版和ONNX版(推荐)。
☆54Oct 12, 2024Updated last year
Alternatives and similar repositories for My-FunASR
Users that are interested in My-FunASR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 介绍docker、docker compose的使用。☆20Sep 4, 2024Updated 2 years ago
- Lattice combination algorithm to combine inaccurate transcripts with hypothesis lattices☆16Mar 19, 2024Updated 2 years ago
- Decoders from Kaldi using OpenFst☆35Apr 10, 2026Updated 5 months ago
- Experiment with JNI access to some Kaldi functions.☆12Dec 31, 2018Updated 7 years ago
- A ctc decoder for both online and offline asr model☆66Nov 18, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- E2E system with LF-MMI; word N-gram for Mandarin☆167Apr 29, 2022Updated 4 years ago
- Implementation of the contextual biasing for ASR decoding on GPUs without lattice generation. The code supports submission to Interspeech…☆21Sep 25, 2023Updated 2 years ago
- C/C++实现Python音频处理库librosa中melspectrogram的计算过程☆31Jan 14, 2022Updated 4 years ago
- Paper, Code and Resources for Speech Language Model and End2End Speech Dialogue System.☆204Jun 7, 2026Updated 3 months ago
- (WIP)long form speech generatoins☆30Apr 2, 2025Updated last year
- ASR教程: https://dataxujing.github.io/ASR-paper/☆26Jul 1, 2024Updated 2 years ago
- c# library for decoding CTTransformer punc models, which can add punctuation to Chinese and English texts☆14Aug 18, 2025Updated last year
- Went online decode demo☆31Apr 28, 2021Updated 5 years ago
- via->yolo, yolo->via☆16Aug 4, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Conversion of recurrent neural network language models to weighted finite state transducers☆58Jun 1, 2018Updated 8 years ago
- Production First and Production Ready End-to-End Keyword Spotting Toolkit☆12May 30, 2022Updated 4 years ago
- Fine-Tune Whisper with Transformers and PEFT☆58Nov 4, 2023Updated 2 years ago
- We Speech Transcript based on LLM, in 300 lines of code.☆182Jun 20, 2025Updated last year
- FunAudioLLM homepage☆17Dec 11, 2024Updated last year
- CTC decoder with hotwords for ASR.☆41Sep 9, 2026Updated last week
- One command to build TLG.fst for WeNet.☆30Oct 11, 2022Updated 3 years ago
- ☆51Nov 26, 2023Updated 2 years ago
- One command to start a streaming ASR server.☆12Oct 2, 2024Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- 语音识别模型pytorch转ONNX转MNN,C++实现部署☆84Sep 1, 2022Updated 4 years ago
- MelGAN and Tacotron 2 in PyTorch☆11Oct 22, 2019Updated 6 years ago
- This is an extension of kaldi speech recognition software which allows to perform decoding of speech with hybrid word and phoneme graphs.…☆11Feb 4, 2020Updated 6 years ago
- ASR client for Triton ASR Service☆39Jan 12, 2026Updated 8 months ago
- ☆23Jul 8, 2019Updated 7 years ago
- ☆28Apr 24, 2026Updated 4 months ago
- 从Kaldi中裁剪的轻量级语音识别解码推理框架,目前实现了MFCC+GMM+Viterbi,不依赖OpenFST、OpenBLAS等库☆22Jul 31, 2021Updated 5 years ago
- ☆12Jul 11, 2024Updated 2 years ago
- ☆31Aug 8, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Backpropagation Latent Dirichlet Allocation (a third-party reimplementation of paper "End-to-end Learning of LDA by Mirror-Descent Back P…☆15Oct 5, 2016Updated 9 years ago
- STT Service based on Kaldi ASR☆15Aug 17, 2018Updated 8 years ago
- Code associated with the paper: CTC-DRO: Robust Optimization for Reducing Language Disparities in Speech Recognition.☆17May 16, 2025Updated last year
- Text-to-dysarthric speech (TTDS) synthesis. An implementation using the Grad-TTS model with the TORGO database.☆15Mar 15, 2025Updated last year
- uyghur text resource crawled from website☆12Dec 25, 2015Updated 10 years ago
- funasr语音转文字的简单api版本,funasr+fastapi,方便部署在服务器上☆13Aug 10, 2024Updated 2 years ago
- ASR (Automatic Speech Recognition) for real-time streamed audio powered by Whisper and tranformers☆36Apr 22, 2026Updated 4 months ago