A demo-level low-latency, high-throughput inference engine for whisper
☆20Nov 9, 2025Updated 10 months ago
Alternatives and similar repositories for nano-whisper
Users that are interested in nano-whisper are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- wav2vec2 audio classification for prosodic boundary detection and other tasks☆42Aug 11, 2023Updated 3 years ago
- A high-performance batch audio transcription tool using nvidia/parakeet-tdt-0.6b-v2 to generate accurate, well-segmented SRT subtitles, w…☆18Dec 9, 2025Updated 9 months ago
- Compute WER and SER for speech recognition evaluation☆28Jun 6, 2026Updated 3 months ago
- ☆10Sep 25, 2024Updated last year
- Multispeaker Community Vocoder Model for DiffSinger☆39Aug 11, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Code for InterSpeech 2024 Paper: LipGER: Visually-Conditioned Generative Error Correction for Robust Automatic Speech Recognition☆19Jul 16, 2024Updated 2 years ago
- Chinese-Mimi 是对 Moshi 模型的声码器进行了中文语料上的适配。☆39Mar 13, 2025Updated last year
- Zeta implementation of a reusable and plug in and play feedforward from the paper "Exponentially Faster Language Modeling"☆16Nov 11, 2024Updated last year
- ☆16Jun 25, 2024Updated 2 years ago
- The case study and multilingfual performance of ICASSP submission☆24Sep 24, 2022Updated 3 years ago
- SLT 2024 Challenge: Post-ASR-Speaker-Tagging☆16Jun 16, 2024Updated 2 years ago
- 一款隐身于 Mac 摄像头下方的智能提词器:专为视频录制、直播与会议设计,帮您保持自然眼神交流。支持苹果自带语音识别与本地 AI 大模型,能随着您的真实语速自动跟踪和滚动文案,彻底告别忘词与手动滑屏的烦恼。☆19Feb 24, 2026Updated 6 months ago
- Port of Funasr's Paraformer model in C/C++☆43Jun 19, 2024Updated 2 years ago
- Serving Next Generation Experimental Tracking for Machine Learning Operations☆15Mar 5, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Running LLaMA 3 with Rust.☆10May 21, 2024Updated 2 years ago
- A Rust crate offering similar functionality to the Python transformers package using Candle.☆15Nov 19, 2024Updated last year
- Script to generate VAD dataset used in Asteroid recipe☆21Sep 30, 2021Updated 4 years ago
- ☆12Jun 14, 2024Updated 2 years ago
- ☆24Aug 1, 2026Updated last month
- Sampling techniques for Candle.☆21Apr 3, 2024Updated 2 years ago
- This is a project of Interspeech2021 paper "SpecMix : A Mixed Sample Data Augmentation method for Training with Time-Frequency Domain Fea…☆11Sep 27, 2022Updated 3 years ago
- Script to demonstrate how to use a Language Model for Semantic Turn Detection. Refer to blog post for full details.☆19May 9, 2025Updated last year
- 适用于 diffsinger 的多功能工具集☆10Apr 2, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- The official repository for AdaMuon☆39Aug 27, 2025Updated last year
- NISQA - Non-Intrusive Speech Quality and TTS Naturalness Assessment☆16Apr 13, 2022Updated 4 years ago
- ☆12Sep 12, 2024Updated 2 years ago
- A ctypes wrapper around the BASS audio library by un4seen.com☆11May 1, 2024Updated 2 years ago
- A chinese singing voice dataset, professional male singer, 105 songs, 132 minutes☆13Oct 19, 2023Updated 2 years ago
- Digital Speech Processing in PyTorch.☆15Aug 12, 2022Updated 4 years ago
- The source code for the paper CrossSinger (asru2023)☆18Oct 12, 2023Updated 2 years ago
- The Bytepiece Tokenizer Implemented in Rust.☆15Nov 28, 2023Updated 2 years ago
- Automatically Update LLM Papers Daily using Github Actions. Ref: https://github.com/Vincentqyw/cv-arxiv-daily☆10Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- This is the accompanying repository to the paper - Automatic Estimation of Singing Voice Musical Dynamics☆16Oct 28, 2024Updated last year
- We Speech Toolkit, LLM based Speech Toolkit for Speech Understanding, Generation, and Interaction☆212Jul 17, 2026Updated 2 months ago
- faster inference☆27Jan 20, 2025Updated last year
- A Feishu/Lark AI agent bot☆16Feb 27, 2026Updated 6 months ago
- ☆20Oct 5, 2025Updated 11 months ago
- 👂 Typing is slow, talk to me. The project name means ' i am tired ' in Chinese (我累了). This is a AI efficiency assistant, complete your d…☆16Jun 8, 2024Updated 2 years ago
- ☆11Nov 2, 2024Updated last year