ASR (Automatic Speech Recognition) for real-time streamed audio powered by Whisper and tranformers
☆36Apr 22, 2026Updated 2 months ago
Alternatives and similar repositories for realtime-whisper
Users that are interested in realtime-whisper are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- funasr语音转文字的简单api版本,funasr+fastapi,方便部署在服务器上☆13Aug 10, 2024Updated last year
- ☆11May 7, 2022Updated 4 years ago
- Data manipulation and transformation for audio signal processing, powered by PyTorch☆10Sep 30, 2024Updated last year
- One command to start a streaming ASR server.☆12Oct 2, 2024Updated last year
- Code for SLT 2016 paper on Grapheme-to-Phoneme conversion using attention based encoder-decoder models☆15Feb 20, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- sherpa with mlx☆15Aug 2, 2025Updated 11 months ago
- ☆12Jul 11, 2024Updated 2 years ago
- ☆52Nov 26, 2023Updated 2 years ago
- 北京航空航天大学课程资料共享仓库☆10Apr 21, 2019Updated 7 years ago
- ASR_LLM_TTS前端项目☆15Dec 3, 2024Updated last year
- Code associated with the paper: CTC-DRO: Robust Optimization for Reducing Language Disparities in Speech Recognition.☆17May 16, 2025Updated last year
- This repository provides a Docker image for CosyVoice☆27Dec 22, 2024Updated last year
- <综合> Funasr语音识别,调用Qwen大模型回答,通过GPTSovits输出语音的ai程序,其中调用模型还是在线,后续将添加离线大模型☆13Nov 30, 2024Updated last year
- Train no-reference speech quality estimators with multiple datasets via learned, per-dataset alignments.☆18Aug 1, 2025Updated 11 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆16Nov 9, 2023Updated 2 years ago
- ☆15Oct 19, 2024Updated last year
- Evaluation Metrics Used For The Performance Evaluation of Voice Conversion (VC) Models☆19Jul 8, 2025Updated last year
- Colab notebooks for Next-gen Kaldi☆31Oct 12, 2025Updated 9 months ago
- Simple voice activity detection (VAD) algorithm in Python☆15Aug 10, 2023Updated 2 years ago
- API to load and query documents using RAG☆14Sep 25, 2023Updated 2 years ago
- PitchVC: Pitch Conditioned Any-to-Many Voice Conversion☆35Jun 6, 2024Updated 2 years ago
- paraformer(chinense asr) online onnx runtime for python☆54Mar 27, 2024Updated 2 years ago
- 简单实现VAD+声纹锁+SenseVoice完成类语音实时转录的小项目☆42Sep 23, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- FunASR安卓端侧离线版本2pass全模式☆15Sep 4, 2023Updated 2 years ago
- A streaming whisper server for on-prem transcription☆23Aug 15, 2024Updated last year
- A Docker image with Llama Index, Lang Chain, and a few other popular AI packages installed by default☆11Nov 19, 2025Updated 8 months ago
- Conformer RNN-Transducer☆14May 25, 2022Updated 4 years ago
- 基于FunASR实现语音识别,包含常规版和ONNX版(推荐)。☆53Oct 12, 2024Updated last year
- FreeSWITCH ASR module fork from mod_audio_stream, use FunASR online cpu version☆20Jun 27, 2025Updated last year
- 修复funasr中seaco-paraformer导出onnx后没有时间戳的bug☆25Sep 12, 2024Updated last year
- (WIP)long form speech generatoins☆30Apr 2, 2025Updated last year
- ☆33Feb 4, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Adding a multi-text multi-speaker script (diffe) that is based on a script from asiff00 on issue 61 for Sesame: A Conversational Speech G…☆26Mar 28, 2025Updated last year
- funasr-android 本地化部署 生成.so给安卓那边调用☆50Mar 20, 2025Updated last year
- Any-to-one voice conversion using the data augment strategy: pitch shifted and duration remained.☆34Jan 10, 2022Updated 4 years ago
- Informed Rapidly-exploring Random Tree-Star with C# Programming☆10Nov 6, 2021Updated 4 years ago
- Compute WER and SER for speech recognition evaluation☆27Jun 6, 2026Updated last month
- This is a web-based intelligent dialogue program built using ASR, LLM, and TTS.☆25Dec 3, 2024Updated last year
- Whisper realtime streaming for long speech-to-text transcription and translation☆22Nov 4, 2024Updated last year