基于FunASR实现语音识别,包含常规版和ONNX版(推荐)。
☆54Oct 12, 2024Updated last year
Alternatives and similar repositories for My-FunASR
Users that are interested in My-FunASR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 介绍docker、docker compose的使用。☆20Sep 4, 2024Updated 2 years ago
- Decoders from Kaldi using OpenFst☆35Apr 10, 2026Updated 6 months ago
- Experiment with JNI access to some Kaldi functions.☆12Dec 31, 2018Updated 7 years ago
- A ctc decoder for both online and offline asr model☆66Nov 18, 2023Updated 2 years ago
- E2E system with LF-MMI; word N-gram for Mandarin☆167Apr 29, 2022Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- magicspeech competition recipe☆18Jun 29, 2020Updated 6 years ago
- Implementation of the contextual biasing for ASR decoding on GPUs without lattice generation. The code supports submission to Interspeech…☆21Sep 25, 2023Updated 3 years ago
- Paper, Code and Resources for Speech Language Model and End2End Speech Dialogue System.☆205Jun 7, 2026Updated 4 months ago
- (WIP)long form speech generatoins☆30Apr 2, 2025Updated last year
- ASR教程: https://dataxujing.github.io/ASR-paper/☆26Jul 1, 2024Updated 2 years ago
- c# library for decoding CTTransformer punc models, which can add punctuation to Chinese and English texts☆14Aug 18, 2025Updated last year
- Went online decode demo☆31Apr 28, 2021Updated 5 years ago
- 使用Qwen1.5-0.5B-Chat模型进行通用信息抽取任务的微调,旨在: 验证生成式方法相较于抽取式NER的效果; 为新手提供简易的模型微调流程,尽量减少代码量; 大模型训练的数据格式处理。☆14Sep 6, 2024Updated 2 years ago
- some ncnn demos of FunASR☆29Sep 23, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Production First and Production Ready End-to-End Keyword Spotting Toolkit☆12May 30, 2022Updated 4 years ago
- Data manipulation and transformation for audio signal processing, powered by PyTorch☆10Sep 30, 2024Updated 2 years ago
- Fine-Tune Whisper with Transformers and PEFT☆58Nov 4, 2023Updated 2 years ago
- We Speech Transcript based on LLM, in 300 lines of code.☆182Jun 20, 2025Updated last year
- FunAudioLLM homepage☆17Dec 11, 2024Updated last year
- CTC decoder with hotwords for ASR.☆41Sep 9, 2026Updated last month
- One command to build TLG.fst for WeNet.☆30Oct 11, 2022Updated 3 years ago
- ☆15Mar 15, 2022Updated 4 years ago
- One command to start a streaming ASR server.☆12Oct 2, 2024Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- 语音识别模型pytorch转ONNX转MNN,C++实现部署☆84Sep 1, 2022Updated 4 years ago
- MelGAN and Tacotron 2 in PyTorch☆11Oct 22, 2019Updated 6 years ago
- This is an extension of kaldi speech recognition software which allows to perform decoding of speech with hybrid word and phoneme graphs.…☆11Feb 4, 2020Updated 6 years ago
- ASR client for Triton ASR Service☆39Jan 12, 2026Updated 8 months ago
- ☆28Apr 24, 2026Updated 5 months ago
- 从Kaldi中裁剪的轻量级语音识别解码推理框架,目前实现了MFCC+GMM+Viterbi,不依赖OpenFST、OpenBLAS等库☆22Jul 31, 2021Updated 5 years ago
- A fast parallel PyTorch implementation of the "CIF: Continuous Integrate-and-Fire for End-to-End Speech Recognition" https://arxiv.org/ab…☆37Feb 10, 2024Updated 2 years ago
- sherpa with mlx☆15Aug 2, 2025Updated last year
- Backpropagation Latent Dirichlet Allocation (a third-party reimplementation of paper "End-to-end Learning of LDA by Mirror-Descent Back P…☆15Oct 5, 2016Updated 10 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- STT Service based on Kaldi ASR☆15Aug 17, 2018Updated 8 years ago
- Code associated with the paper: CTC-DRO: Robust Optimization for Reducing Language Disparities in Speech Recognition.☆17May 16, 2025Updated last year
- PaliGemma Inference and Fine Tuning☆13May 15, 2024Updated 2 years ago
- Text-to-dysarthric speech (TTDS) synthesis. An implementation using the Grad-TTS model with the TORGO database.☆16Mar 15, 2025Updated last year
- uyghur text resource crawled from website☆12Dec 25, 2015Updated 10 years ago
- funasr语音转文字的简单api版本,funasr+fastapi,方便部署在服务器上☆13Aug 10, 2024Updated 2 years ago
- ASR (Automatic Speech Recognition) for real-time streamed audio powered by Whisper and tranformers☆36Updated this week