Efficient approach to speaker diarization using voice characteristics extraction
☆110Jun 26, 2026Updated 2 months ago
Alternatives and similar repositories for WhoSpeaks
Users that are interested in WhoSpeaks are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Real-time speaker diarization using straightforward, intuitive logic - High accuracy thanks to SpeechBrain/Pyannote-WeSpeaker models☆33May 7, 2026Updated 3 months ago
- Command Your World with Voice☆813Jun 17, 2025Updated last year
- Simulates talk with an AI that can express emotions☆91Apr 4, 2026Updated 4 months ago
- A python package to build AI-powered real-time audio applications☆2,023Jun 19, 2026Updated 2 months ago
- Provide Gradio custom components to make the diarization-based audio labeling process easier and faster.☆72Apr 22, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Speaker Diarization with Transformers☆70Jun 8, 2025Updated last year
- Private, real-time speaker diarization on hardware you control. See who is speaking as it happens, no third-party cloud required.☆26Jul 25, 2026Updated last month
- Very fast, accurate speaker diarization☆293Updated this week
- Roomey is a multi-purpose Voice Agent designed to run your personal and business life.☆67Jun 15, 2025Updated last year
- Low latency ai companion voice talk in 60 lines of code using faster_whisper and elevenlabs input streaming☆322Jun 17, 2025Updated last year
- DiariZen Explained: A Tutorial for the Open Source State-of-the-Art Speaker Diarization Pipeline.☆24Aug 10, 2026Updated 3 weeks ago
- Real-time processing and delivery of sentences from a continuous stream of characters or text chunks.☆88Aug 25, 2026Updated last week
- Converts text to speech in realtime☆4,020Updated this week
- ☆12Sep 28, 2024Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Cog implementation of transcribing + diarization pipeline with Whisper & Pyannote☆237Jun 11, 2026Updated 2 months ago
- PAFTS : Library That Preprocessing Audio For TTS.☆27Nov 15, 2024Updated last year
- Blind Source Separation and Dereverberation☆21Mar 26, 2021Updated 5 years ago
- auto fine tune of models with synthetic data☆78Feb 14, 2024Updated 2 years ago
- This repo contains script to download MUSIC dataset from youtube☆12Jan 19, 2024Updated 2 years ago
- An application-layer router for Skupper networks☆21Aug 5, 2026Updated 3 weeks ago
- Multi-Stage Face-Voice Association Learning with Keynote Speaker Diarization (ACM MM 2024)☆22Jul 25, 2024Updated 2 years ago
- A python package for deep multilingual punctuation prediction.☆169Aug 21, 2024Updated 2 years ago
- audio, NLP, ML with huggingface, nvidia/nemo, speechbrain☆11Sep 4, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Local AI talk with a custom voice based on Zephyr 7B model. Uses RealtimeSTT with faster_whisper for transcription and RealtimeTTS with C…☆731Jun 17, 2025Updated last year
- Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper☆5,633Aug 15, 2026Updated 2 weeks ago
- 🎹 pyannote + 🗒 notebook = pyannotebook☆27Jun 12, 2023Updated 3 years ago
- [Colab Demo Code] OneFormer: One Transformer to Rule Universal Image Segmentation.☆14May 24, 2023Updated 3 years ago
- Python Wrapper around Ollama API Endpoints☆12Jan 26, 2024Updated 2 years ago
- C++ version of pyannote audio speaker diarizaiton pipeline☆22Feb 14, 2024Updated 2 years ago
- Latent Large Language Models☆19Aug 24, 2024Updated 2 years ago
- Authenticating proxy server for connecting to 3rd party APIs☆17Updated this week
- A highly-customizable OpenAI gym environment to train & evaluate RL agents trading stocks and crypto.☆21Jun 6, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Example Cloudflare Workers project showing how to return HTML responses with enriched region data☆14May 3, 2021Updated 5 years ago
- Some comprehensive papers about speaker diarization☆370Mar 24, 2026Updated 5 months ago
- An unofficial implementation of the Personal VAD speaker-conditioned voice activity detection method. Bachelor's thesis project.☆90Sep 22, 2022Updated 3 years ago
- Speech Diarization for scrum automation☆111Jul 27, 2023Updated 3 years ago
- 💬 ASR FastAPI server using faster-whisper and Multi-Scale Auto-Tuning Spectral Clustering for diarization.☆219Oct 30, 2024Updated last year
- ☆13May 23, 2024Updated 2 years ago
- Official repository for the "Powerset multi-class cross entropy loss for neural speaker diarization" paper published in Interspeech 2023.☆95Oct 18, 2023Updated 2 years ago