A streaming whisper server for on-prem transcription
☆23Aug 15, 2024Updated last year
Alternatives and similar repositories for streaming-whisper-server
Users that are interested in streaming-whisper-server are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Data manipulation and transformation for audio signal processing, powered by PyTorch☆10Sep 30, 2024Updated last year
- One command to start a streaming ASR server.☆12Oct 2, 2024Updated last year
- ☆12Jul 11, 2024Updated 2 years ago
- Code associated with the paper: CTC-DRO: Robust Optimization for Reducing Language Disparities in Speech Recognition.☆17May 16, 2025Updated last year
- funasr语音转文字的简单api版本,funasr+fastapi,方便部署在服务器上☆13Aug 10, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Whisper realtime streaming for long speech-to-text transcription and translation☆121Jan 29, 2024Updated 2 years ago
- 基于wenet的短时在线语音识别服务☆11Feb 25, 2023Updated 3 years ago
- ☆14Aug 9, 2021Updated 4 years ago
- <综合> Funasr语音识别,调用Qwen大模型回答,通过GPTSovits输出语音的ai程序,其中调用模型还是在线,后续将添加离线大模型☆13Nov 30, 2024Updated last year
- Train no-reference speech quality estimators with multiple datasets via learned, per-dataset alignments.☆18Aug 1, 2025Updated 11 months ago
- ☆11Dec 24, 2024Updated last year
- ☆16Nov 9, 2023Updated 2 years ago
- ☆15Oct 19, 2024Updated last year
- Whisper combined with Silero VAD, for improved long-form transcriptions☆55Dec 11, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Simple voice activity detection (VAD) algorithm in Python☆15Aug 10, 2023Updated 2 years ago
- A OCR Project for Reading New and Old Kannada Texts☆10Aug 31, 2024Updated last year
- ☆14Dec 9, 2014Updated 11 years ago
- [DEPRECATED] Baseline Project for Semantic Searching☆10Oct 15, 2018Updated 7 years ago
- Column Networks for Collective Classification: A novel deep learning model for collective classification in multi-relational domains☆12Nov 22, 2016Updated 9 years ago
- Sthaan uses AI to create digital addresses with local language support in voice/text, making it easier for people to find and reach locat…☆12Nov 17, 2024Updated last year
- A C# abstraction on top of Godot's input events that makes life just a little bit easier☆10Feb 1, 2018Updated 8 years ago
- Android app for the Hole in your Palm project, making LLMs accessible on-device!☆19May 3, 2024Updated 2 years ago
- This repository contains the code for the paper "Botometer 101: Social bot practicum for computational social scientists."☆11Oct 6, 2022Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- 🎯 Speech Recognition Challenge by Speech Lab - IIT Madras☆10Nov 5, 2020Updated 5 years ago
- FunASR安卓端侧离线版本2pass全模式☆15Sep 4, 2023Updated 2 years ago
- This is a project focused on Faster Whisper, a streaming speech recognition project.☆18Sep 27, 2024Updated last year
- Towards Building Text-To-Speech Systems for the Next Billion Users - Microsoft Research Intern Work - Accepted at ICASSP 2023☆57May 7, 2023Updated 3 years ago
- Conformer RNN-Transducer☆14May 25, 2022Updated 4 years ago
- WebRTC-based real-time audio streaming with Faster Whisper ASR integration for live speech-to-text transcription.☆13Sep 27, 2024Updated last year
- Use LoRA technique to improve training Large Language Model☆13Jul 25, 2023Updated 3 years ago
- FreeSWITCH ASR module fork from mod_audio_stream, use FunASR online cpu version☆20Jun 27, 2025Updated last year
- 修复funasr中seaco-paraformer导出onnx后没有时间戳的bug☆25Sep 12, 2024Updated last year
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- FSDS Webinar 1: Real-Time Machine Learning Inference with Spark Streaming and Kafka☆10Feb 17, 2025Updated last year
- (WIP)long form speech generatoins☆30Apr 2, 2025Updated last year
- Intrusion. Custom Asterisk dial plan for listen, whisper and barge in calls. For Asterisk FreePBX, Issabel, Asterisk based Elastix call c…☆16Jul 9, 2021Updated 5 years ago
- ☆33Feb 4, 2025Updated last year
- ☆10Sep 10, 2023Updated 2 years ago
- ☆15Dec 8, 2022Updated 3 years ago
- Fine-Tune Whisper with Transformers and PEFT☆58Nov 4, 2023Updated 2 years ago