Efficient approach to speaker diarization using voice characteristics extraction
☆109Jun 26, 2026Updated 3 weeks ago
Alternatives and similar repositories for WhoSpeaks
Users that are interested in WhoSpeaks are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Real-time speaker diarization using straightforward, intuitive logic - High accuracy thanks to SpeechBrain/Pyannote-WeSpeaker models☆30May 7, 2026Updated 2 months ago
- Command Your World with Voice☆811Jun 17, 2025Updated last year
- Tr-VAD: An Efficient Transformer based Voice Activity Detection Model☆18Aug 1, 2024Updated last year
- Simulates talk with an AI that can express emotions☆88Apr 4, 2026Updated 3 months ago
- A python package to build AI-powered real-time audio applications☆2,003Jun 19, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Transcribe desktop audio/computer audio in real-time and locally (Streaming ASR), using TorchAudio and Emformer-RNNT model for inference,…☆14May 7, 2024Updated 2 years ago
- Provide Gradio custom components to make the diarization-based audio labeling process easier and faster.☆71Apr 22, 2026Updated 2 months ago
- Very fast, accurate speaker diarization☆284Jun 11, 2026Updated last month
- DiariZen Explained: A Tutorial for the Open Source State-of-the-Art Speaker Diarization Pipeline.☆21Apr 24, 2026Updated 2 months ago
- Real-time processing and delivery of sentences from a continuous stream of characters or text chunks.☆82Updated this week
- Converts text to speech in realtime☆3,995May 31, 2026Updated last month
- ☆11Sep 28, 2024Updated last year
- Simple PyTorch Denoisers for Waveform Audio☆41Apr 4, 2026Updated 3 months ago
- Cog implementation of transcribing + diarization pipeline with Whisper & Pyannote☆237Jun 11, 2026Updated last month
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- PAFTS : Library That Preprocessing Audio For TTS.☆27Nov 15, 2024Updated last year
- auto fine tune of models with synthetic data☆78Feb 14, 2024Updated 2 years ago
- Automatically setup the AISHELL-4 and MSDWild dataset for usage with pyannote-database (and pyannote-audio)☆15Oct 22, 2025Updated 8 months ago
- This repo contains script to download MUSIC dataset from youtube☆12Jan 19, 2024Updated 2 years ago
- ☆88Jul 31, 2025Updated 11 months ago
- Multi Browser Kango Extension for BGPView - A DNS and BGP network visualizer☆10May 16, 2017Updated 9 years ago
- Multi-Stage Face-Voice Association Learning with Keynote Speaker Diarization (ACM MM 2024)☆22Jul 25, 2024Updated last year
- audio, NLP, ML with huggingface, nvidia/nemo, speechbrain☆11Sep 4, 2023Updated 2 years ago
- Local AI talk with a custom voice based on Zephyr 7B model. Uses RealtimeSTT with faster_whisper for transcription and RealtimeTTS with C…☆726Jun 17, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆14Aug 22, 2024Updated last year
- Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper☆5,603Feb 23, 2026Updated 4 months ago
- Svelte app to generate audiobooks using XTTS☆12Feb 13, 2024Updated 2 years ago
- Exploring Binary Classification Loss for Speaker Verification☆18Jul 18, 2023Updated 3 years ago
- C++ version of pyannote audio speaker diarizaiton pipeline☆22Feb 14, 2024Updated 2 years ago
- Get started using Deepgram's Live Transcription with this Flask demo app☆46Apr 11, 2026Updated 3 months ago
- A highly-customizable OpenAI gym environment to train & evaluate RL agents trading stocks and crypto.☆20Jun 6, 2023Updated 3 years ago
- Example Cloudflare Workers project showing how to return HTML responses with enriched region data☆14May 3, 2021Updated 5 years ago
- An unofficial implementation of the Personal VAD speaker-conditioned voice activity detection method. Bachelor's thesis project.☆89Sep 22, 2022Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A scalable solution that simplifies the integration of ComfyUI for developers☆11Jul 15, 2024Updated 2 years ago
- Speech Diarization for scrum automation☆111Jul 27, 2023Updated 2 years ago
- Hand MoCap 3d viewer for the IJCV'16 paper "Capturing Hands in Action using Discriminative Salient Points and Physics Simulation"☆11May 19, 2016Updated 10 years ago
- 💬 ASR FastAPI server using faster-whisper and Multi-Scale Auto-Tuning Spectral Clustering for diarization.☆220Oct 30, 2024Updated last year
- ☆16Feb 19, 2026Updated 5 months ago
- ☆13May 23, 2024Updated 2 years ago
- ☆19Jul 1, 2023Updated 3 years ago