基于FunASR实现语音识别,包含常规版和ONNX版(推荐)。
☆53Oct 12, 2024Updated last year
Alternatives and similar repositories for My-FunASR
Users that are interested in My-FunASR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Lattice combination algorithm to combine inaccurate transcripts with hypothesis lattices☆16Mar 19, 2024Updated 2 years ago
- Decoders from Kaldi using OpenFst☆35Apr 10, 2026Updated 4 months ago
- Experiment with JNI access to some Kaldi functions.☆12Dec 31, 2018Updated 7 years ago
- magicspeech competition recipe☆18Jun 29, 2020Updated 6 years ago
- Implementation of the contextual biasing for ASR decoding on GPUs without lattice generation. The code supports submission to Interspeech…☆21Sep 25, 2023Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Paper, Code and Resources for Speech Language Model and End2End Speech Dialogue System.☆202Jun 7, 2026Updated 2 months ago
- (WIP)long form speech generatoins☆30Apr 2, 2025Updated last year
- Computes the MWER (minimum WER) Loss with CTC beam search. Knowledge distillation for CTC loss.☆60Sep 6, 2023Updated 2 years ago
- ASR教程: https://dataxujing.github.io/ASR-paper/☆25Jul 1, 2024Updated 2 years ago
- c# library for decoding CTTransformer punc models, which can add punctuation to Chinese and English texts☆14Aug 18, 2025Updated 11 months ago
- Went online decode demo☆31Apr 28, 2021Updated 5 years ago
- via->yolo, yolo->via☆16Aug 4, 2025Updated last year
- some ncnn demos of FunASR☆29Sep 23, 2024Updated last year
- Production First and Production Ready End-to-End Keyword Spotting Toolkit☆12May 30, 2022Updated 4 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Data manipulation and transformation for audio signal processing, powered by PyTorch☆10Sep 30, 2024Updated last year
- Fine-Tune Whisper with Transformers and PEFT☆58Nov 4, 2023Updated 2 years ago
- We Speech Transcript based on LLM, in 300 lines of code.☆182Jun 20, 2025Updated last year
- FunAudioLLM homepage☆17Dec 11, 2024Updated last year
- CTC decoder with hotwords for ASR.☆39Updated this week
- One command to build TLG.fst for WeNet.☆30Oct 11, 2022Updated 3 years ago
- ☆52Nov 26, 2023Updated 2 years ago
- One command to start a streaming ASR server.☆12Oct 2, 2024Updated last year
- 语音识别模型pytorch转ONNX转MNN,C++实现部署☆85Sep 1, 2022Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- MelGAN and Tacotron 2 in PyTorch☆11Oct 22, 2019Updated 6 years ago
- This is an extension of kaldi speech recognition software which allows to perform decoding of speech with hybrid word and phoneme graphs.…☆11Feb 4, 2020Updated 6 years ago
- ☆28Apr 24, 2026Updated 3 months ago
- 从Kaldi中裁剪的轻量级语音识别解码推理框架,目前实现了MFCC+GMM+Viterbi,不依赖OpenFST、OpenBLAS等库☆22Jul 31, 2021Updated 5 years ago
- A fast parallel PyTorch implementation of the "CIF: Continuous Integrate-and-Fire for End-to-End Speech Recognition" https://arxiv.org/ab…☆37Feb 10, 2024Updated 2 years ago
- ☆29Aug 8, 2024Updated 2 years ago
- sherpa with mlx☆15Aug 2, 2025Updated last year
- ☆12Jul 11, 2024Updated 2 years ago
- Backpropagation Latent Dirichlet Allocation (a third-party reimplementation of paper "End-to-end Learning of LDA by Mirror-Descent Back P…☆15Oct 5, 2016Updated 9 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- STT Service based on Kaldi ASR☆15Aug 17, 2018Updated 7 years ago
- PaliGemma Inference and Fine Tuning☆13May 15, 2024Updated 2 years ago
- Text-to-dysarthric speech (TTDS) synthesis. An implementation using the Grad-TTS model with the TORGO database.☆14Mar 15, 2025Updated last year
- uyghur text resource crawled from website☆12Dec 25, 2015Updated 10 years ago
- funasr语音转文字的简单api版本,funasr+fastapi,方便部署在服务器上☆13Aug 10, 2024Updated 2 years ago
- ASR (Automatic Speech Recognition) for real-time streamed audio powered by Whisper and tranformers☆36Apr 22, 2026Updated 3 months ago
- The repo contains our code of ``Semantic Mask for Transformer based End-to-End Speech Recognition"☆39Jun 9, 2020Updated 6 years ago