Whisper realtime streaming for long speech-to-text transcription and translation
☆22Nov 4, 2024Updated last year
Alternatives and similar repositories for whisper_streaming
Users that are interested in whisper_streaming are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- One command to start a streaming ASR server.☆12Oct 2, 2024Updated last year
- Authenticating proxy server for connecting to 3rd party APIs☆17Updated this week
- 基于 Sherpa-ONNX 实现在线下载模型的端 侧实时语音识别应用(Implement speech recognition based on Sherpa-ONNX by downloading the model online.)☆32Feb 27, 2025Updated last year
- Official data release for FaceMap, to present in Siggraph Asia 2024☆13Nov 1, 2024Updated last year
- sherpa with mlx☆15Aug 2, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆12Jul 11, 2024Updated 2 years ago
- The repo for: TriHuman: A Real-time and Controllable Tri-plane Representation for Detailed Human Geometry and Appearance Synthesis☆19Nov 15, 2025Updated 9 months ago
- InfNeRF: Towards Infinite Scale NeRF Rendering with O(log n) Space Complexity☆12Jan 3, 2026Updated 8 months ago
- Code associated with the paper: CTC-DRO: Robust Optimization for Reducing Language Disparities in Speech Recognition.☆17May 16, 2025Updated last year
- 基于wenet的短时在线语音识别服务☆11Feb 25, 2023Updated 3 years ago
- ☆14Aug 9, 2021Updated 5 years ago
- <综合> Funasr语音识别,调用Qwen大模型回答,通过GPTSovits输出语音的ai程序,其中调用模型还是在线,后续将添加离线大模型☆13Nov 30, 2024Updated last year
- ASR_LLM_TTS前端项目☆15Dec 3, 2024Updated last year
- ☆24Jan 22, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- X-SLAM: Scalable Dense SLAM for Task-aware Optimization using CSFD (ACM SIGGRAPH 2024)☆15Jul 24, 2024Updated 2 years ago
- HAAR: Text-Conditioned Generative Model of 3D Strand-based Human Hairstyles (CVPR 2024)☆92Oct 9, 2024Updated last year
- Asset management solution☆12Mar 2, 2023Updated 3 years ago
- Train no-reference speech quality estimators with multiple datasets via learned, per-dataset alignments.☆19Aug 1, 2025Updated last year
- ☆11Dec 24, 2024Updated last year
- ☆16Nov 9, 2023Updated 2 years ago
- ☆15Jun 21, 2023Updated 3 years ago
- The Receipt Pattern — auditable AI agent actions in 10 lines. Every agent action leaves a receipt.☆21Mar 11, 2026Updated 5 months ago
- ☆14Sep 21, 2025Updated 11 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Cua.kira is a self-hosted AI desktop agent that automates computer tasks through natural language commands, operating within a containeri…☆19Nov 1, 2025Updated 10 months ago
- Simple voice activity detection (VAD) algorithm in Python☆15Aug 10, 2023Updated 3 years ago
- A GUI to inspect the NeRSemble dataset☆14Apr 11, 2025Updated last year
- ☆50Jan 20, 2025Updated last year
- 📦 Easy Python to Fast Executables☆33Aug 15, 2026Updated 3 weeks ago
- ☆15Jul 2, 2025Updated last year
- [ICLR 2025] Official implementation of "Perm: A Parametric Representation for Multi-Style 3D Hair Modeling"☆123Jul 11, 2026Updated last month
- Wait for async tasks☆13Dec 22, 2022Updated 3 years ago
- ☆24Aug 19, 2026Updated 2 weeks ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆22Apr 25, 2022Updated 4 years ago
- Your Army of GPT-4 Powered Coding Buddies (Boost Your Productivity)☆15Jul 4, 2023Updated 3 years ago
- FunASR安卓端侧离线版本2pass全模式☆15Sep 4, 2023Updated 3 years ago
- A streaming whisper server for on-prem transcription☆23Aug 15, 2024Updated 2 years ago
- [ECCV 2024] PanoFree: Tuning-Free Holistic Multi-view Image Generation with Cross-view Self-Guidance☆24Jul 25, 2024Updated 2 years ago
- This is a project focused on Faster Whisper, a streaming speech recognition project.☆18Sep 27, 2024Updated last year
- The official implementation of our ICCV 2025 paper, "Bring Your Rear Cameras for Egocentric 3D Human Pose Estimation".☆17Dec 3, 2025Updated 9 months ago