Speechlib is a library that unifies speaker diarization, transcription and speaker recognition in a single pipeline to create transcripts for audio conversations with actual speaker names and time tags. This library also contains audio preprocessor functions.
☆266Apr 19, 2026Updated 3 months ago
Alternatives and similar repositories for speechlib
Users that are interested in speechlib are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper☆5,621Updated this week
- Skribify is a powerful transcription and summarization tool that leverages the power of OpenAI's GPT-4 and WhisperAI to generate concise …☆12Aug 6, 2026Updated last week
- A testing repo to share code and thoughts on diarisation☆58Mar 26, 2024Updated 2 years ago
- ☆490Sep 10, 2025Updated 11 months ago
- Acoustic echo cancelation(AEC) is a main algorithm in the pipe line of acoustic devices with KWS or ASR. FNLMS is used.☆19Apr 22, 2019Updated 7 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- On-device speaker diarization powered by deep learning☆77Updated this week
- 💬 ASR FastAPI server using faster-whisper and Multi-Scale Auto-Tuning Spectral Clustering for diarization.☆219Oct 30, 2024Updated last year
- The WhisperX API is a containerized solution for transcribing audio files using the powerful `whisperx` model. This API provides an easy-…☆18Aug 24, 2023Updated 2 years ago
- Verbatim Automatic Speech Recognition with improved word-level timestamps and filler detection☆1,297Aug 6, 2026Updated last week
- This repository contains prompts & best practices to annotate audio clips with a very high degree of details using Audio-Language-Models☆35Oct 13, 2024Updated last year
- Speaker diarization service☆28Jul 2, 2026Updated last month
- ez audio transcription tool with flexible processing and post-processing options☆171Feb 1, 2024Updated 2 years ago
- Docker image for WhisperX by Max Bain☆13Sep 24, 2025Updated 10 months ago
- sherpa with mlx☆15Aug 2, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Neural building blocks for speaker diarization: speech activity detection, speaker change detection, overlapped speech detection, speaker…☆10,417Aug 4, 2026Updated last week
- Official repository of the work "Low-complexity Unsupervised Audio Anomaly Detection exploiting Separable Convolutions and Angular Loss" …☆11Nov 6, 2024Updated last year
- An open-source Kazakh Emotional Text-to-Speech Dataset☆36Aug 1, 2025Updated last year
- Estimating the Age, Height, and Gender of a speaker with their speech signal.☆15Sep 19, 2022Updated 3 years ago
- WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)☆23,570Jul 13, 2026Updated last month
- The official Pytorch implementation of "Frame-wise streaming end-to-end speaker diarization with non-autoregressive self-attention-based …☆187May 7, 2026Updated 3 months ago
- ☆677Sep 24, 2025Updated 10 months ago
- Synchronize Whisper's timestamps over an existing accurate transcription☆165May 28, 2024Updated 2 years ago
- ☆328Jun 14, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Experimental code: sound file preprocessing to optimize Whisper transcriptions without hallucinated texts☆351Nov 12, 2024Updated last year
- Open TTS models, built for streaming on the edge☆45Mar 16, 2025Updated last year
- Automatic Speech Recognition (ASR) system for the Samrómur speech corpus using Kaldi☆12Sep 30, 2022Updated 3 years ago
- A collection of custom tools and extensions for Open WebUI that enhance its capabilities☆12Dec 11, 2024Updated last year
- Minimal extension of OpenAI's Whisper adding speaker diarization with special tokens☆551Nov 6, 2023Updated 2 years ago
- A simple Python + Tkinter + Tesseract-based GUI image-to-text copypaste pad application☆10Sep 14, 2023Updated 2 years ago
- [APSIPA'22] Exploring Speaker Age Estimation on Different Self-Supervised Learning Models☆14Oct 19, 2022Updated 3 years ago
- ☆358Mar 17, 2024Updated 2 years ago
- How to use OpenAIs Whisper to transcribe and diarize audio files☆378Oct 12, 2022Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A nearly-live implementation of OpenAI's Whisper.☆4,227Aug 4, 2026Updated last week
- turnkey self-hosted offline transcription and diarization service with llm summary☆945Jan 18, 2026Updated 6 months ago
- This script is an automated survey bot that conducts political discussions over phone calls. It uses Flask, Twilio's Voice API, OpenAI's …☆12Sep 21, 2023Updated 2 years ago
- Vocoder-Free Non-Parallel Conversion of Whispered Speech With Masked Cycle-Consistent Generative Adversarial Networks☆17Aug 18, 2023Updated 2 years ago
- Companion repo for the paper "PixIT: Joint Training of Speaker Diarization and Speech Separation from Real-world Multi-speaker Recordings…☆106Jan 10, 2025Updated last year
- This repository contains audio samples and supplementary materials accompanying publications by the "Speaker, Voice and Language" team at…☆454Aug 12, 2025Updated last year
- Smart assistant in Telegram bot format for transcribing online meetings☆15Apr 1, 2024Updated 2 years ago