基于 faster-whisper 的伪实时语音转写服务
☆241Apr 29, 2025Updated last year
Alternatives and similar repositories for stream-whisper
Users that are interested in stream-whisper are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This is a project focused on Faster Whisper, a streaming speech recognition project.☆18Sep 27, 2024Updated 2 years ago
- c# library for decoding K2 transducer Models,used in speech recognition (ASR)☆13Aug 20, 2025Updated last year
- Whisper realtime streaming for long speech-to-text transcription and translation☆62Apr 9, 2024Updated 2 years ago
- a gradio webui for faster whisper☆276Apr 28, 2023Updated 3 years ago
- Fine-tune the Whisper speech recognition model to support training without timestamp data, training with timestamp data, and training wit…☆319Dec 22, 2025Updated 9 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- API and websocket server for sensevoice. It has inherited some enhanced features, such as VAD detection, real-time streaming recognition,…☆542Oct 23, 2024Updated last year
- Official code for "F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching"☆14Nov 17, 2024Updated last year
- A nearly-live implementation of OpenAI's Whisper.☆4,295Sep 10, 2026Updated 2 weeks ago
- 一个基于 cloudflare worker 的 LLM API 反向代理,支持 OpenAI, Gemini, Groq 等平台。接口兼容 OpenAI API 规范,可以直接使用 OpenAI SDK 调用。☆255Apr 30, 2026Updated 4 months ago
- Faster Whisper transcription with CTranslate2☆25,598Nov 19, 2025Updated 10 months ago
- Deploy Python + FastAPI project on vercel.com☆18Oct 15, 2024Updated last year
- 基于ESP32制作的透明小电视☆20May 14, 2024Updated 2 years ago
- Real time faster whisper gradio☆24Aug 17, 2025Updated last year
- PersonaTalk Hack☆15Jan 10, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Use React & FastAPI to implement whisper-based demo(使用 「React + FastAPI 」实现的web端 「whisper 」语音识别 demo)☆29Apr 15, 2024Updated 2 years ago
- CTC decoder with hotwords for ASR.☆41Sep 9, 2026Updated 2 weeks ago
- Bert-vits2转写和标注独立整合Webui,整合阿里FunAsr,必剪Asr以及Whisper大模型☆181Jul 10, 2024Updated 2 years ago
- ASR (Automatic Speech Recognition) for real-time streamed audio powered by Whisper and tranformers☆36Updated this week
- A simple implementation of real-time output device audio transcription and translation using "faster_whisper" and "pyaudiowpatch".☆22May 6, 2023Updated 3 years ago
- Whisper realtime streaming for long speech-to-text transcription and translation☆3,673Nov 12, 2025Updated 10 months ago
- 🎤 The easiest way to transcribe audio in Swift☆12Dec 10, 2023Updated 2 years ago
- flow mirror models from JZX AI Labs☆43Sep 30, 2024Updated last year
- Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenA…☆20,517Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- 使用 FastAPI、Streamlit本地部署ChatTTS文本转语音模型,并通过 Docker Compose 进行容器化部署。☆27Sep 27, 2024Updated 2 years ago
- Pseudo Streaming SenseVoice with Hotwords☆471Updated this week
- funasr语音转文字的简单api版本,funasr+fastapi,方便部署在服务器上☆13Aug 10, 2024Updated 2 years ago
- 深度学习☆13Feb 16, 2023Updated 3 years ago
- A enterprise-grade Voice Activity Detector from modelscope and funasr.☆142Apr 26, 2023Updated 3 years ago
- 网页版 ChatGPT 包装成 API☆171Oct 28, 2024Updated last year
- Port of Funasr's Sense-voice model in C/C++☆574Dec 19, 2025Updated 9 months ago
- 微软邮箱 API 客户端 - 基于 Cloudflare Workers 的 Web UI☆48Jan 23, 2025Updated last year
- A quick and optimized solution to manage llama based gguf quantized models, download gguf files, retreive messege formatting, add more mo…☆15Jan 13, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Fine-tune the Whisper speech recognition model to support training without timestamp data, training with timestamp data, and training wit…☆1,218May 8, 2026Updated 4 months ago
- 对比测试不同大语言模型(LLM)性能的工具平台,支持DeepSeek API、Ollama本地模型和VLLM本地模型。A simple tools to test multi models and display the time cost.☆29May 7, 2025Updated last year
- ☆24Jan 4, 2024Updated 2 years ago
- 一个入门的cesium系统,具备基本功能和一些空间分析(剖面,通视,可视域)功能,bug多,后期改☆19Jan 7, 2021Updated 5 years ago
- Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio…☆9,395Updated this week
- Compute WER and SER for speech recognition evaluation☆28Jun 6, 2026Updated 3 months ago
- This is my internship project in NetEase game AI Lab: multi-modal virtual human interaction, through the text to predict the virtual huma…☆10Nov 5, 2018Updated 7 years ago