Apply https://github.com/k2-fsa/sherpa-ncnn in live streaming and WebRTC
☆20Apr 16, 2023Updated 3 years ago
Alternatives and similar repositories for srs-k2
Users that are interested in srs-k2 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Python wrapper for kaldi's arpa2fst☆38Aug 27, 2025Updated 11 months ago
- python wrapper for kaldi's native I/O☆27Jan 9, 2025Updated last year
- PyTorch implementation of Continuous Speech Separation☆12Oct 5, 2022Updated 3 years ago
- ☆10Dec 16, 2022Updated 3 years ago
- c# library for decoding K2 transducer Models,used in speech recognition (ASR)☆13Aug 20, 2025Updated 11 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- BurrMill core☆22Nov 2, 2021Updated 4 years ago
- Python wrapper for OpenFST and its extensions from Kaldi. Also support reading/writing ark/scp files☆56Apr 9, 2026Updated 3 months ago
- Some fast-ish algorithms for batch text search in moderate-sized collections, intended for data cleanup☆79Jun 30, 2025Updated last year
- ☆28Apr 24, 2026Updated 3 months ago
- An ASR toolkit with the freedom of topology☆10Dec 18, 2023Updated 2 years ago
- This solution is not good enough, we're researching a better version: https://github.com/winlinvip/vod-translator so we archive this repo…☆21Apr 17, 2024Updated 2 years ago
- ☆15May 8, 2021Updated 5 years ago
- Podcast Summarizer with LLM Technology☆30May 28, 2025Updated last year
- An echo cancellation library for browsers using DTLN-aec☆26Oct 18, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Data and code related to the ICASSP submission "A comparison of methods for OOV-word recognition"☆17Nov 28, 2021Updated 4 years ago
- Source Code for the Paper "UNIFIED KEYWORD SPOTTING AND AUDIO TAGGING ON MOBILE DEVICES WITH TRANSFORMERS"☆24Mar 6, 2023Updated 3 years ago
- A simple package for Guided source separation (GSS)☆134May 20, 2024Updated 2 years ago
- Kaldi-compatible online & offline feature extraction with PyTorch, supporting CUDA, batch processing, chunk processing, and autograd - P…☆215Jul 10, 2026Updated 3 weeks ago
- Small language toolkit for creation, interpolation and pruning of ARPA language models☆92Aug 6, 2022Updated 3 years ago
- Conversion of recurrent neural network language models to weighted finite state transducers☆58Jun 1, 2018Updated 8 years ago
- Audio samples accompanying publications related to DF-Conformer, a speech enhancement model.☆36Jun 23, 2026Updated last month
- Computes the MWER (minimum WER) Loss with CTC beam search. Knowledge distillation for CTC loss.☆59Sep 6, 2023Updated 2 years ago
- Text utilities, including beam search decoding, tokenizing, and more, built for use in Flashlight.☆78Mar 31, 2026Updated 4 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Decoders from Kaldi using OpenFst☆35Apr 10, 2026Updated 3 months ago
- ☆34Jul 23, 2024Updated 2 years ago
- A collection of all our phonemeizers for dataset construction and inference☆30Feb 21, 2025Updated last year
- c++ Kaldi IO lib (static and dynamic).☆25Nov 26, 2018Updated 7 years ago
- Source code for ICASSP2022 "Pseudo Strong labels for large scale weakly supervised audio tagging"☆31Apr 29, 2022Updated 4 years ago
- Filtering and Noise Adding Tool☆29May 27, 2022Updated 4 years ago
- it's ASR decoder and make graph project☆33May 26, 2022Updated 4 years ago
- Provide accurate offline voice-to-text services for VR,AR and Android platforms, such as oculus quest1/2/pro or pico3/4☆26May 21, 2024Updated 2 years ago
- This is a effective VAD(Voice Activity Detection) for iOS & Android. It is port from google webrtc.☆12Jul 13, 2017Updated 9 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Memory efficient transducer loss computation☆70Jun 10, 2022Updated 4 years ago
- node-addon-api for HarmonyOS/HarmonyNext☆13May 19, 2026Updated 2 months ago
- Greedy Adaptive Dictionary (GAD) is a learning algorithm that sets out to find sparse atoms for speech signals.☆11Oct 1, 2018Updated 7 years ago
- ☆14May 25, 2023Updated 3 years ago
- Example implementation of the Alexa Voice Service Integration for AWS IoT Core for Arm Cortex-M series processors.☆12Feb 24, 2021Updated 5 years ago
- Directional sparse filtering for blind speech separation☆11Jun 8, 2021Updated 5 years ago
- Push FreeSWITCH Realtime info to InfluxDB & PostgreSQL☆15Jul 7, 2020Updated 6 years ago