Efficient approach to speaker diarization using voice characteristics extraction
☆109Jun 26, 2026Updated last month
Alternatives and similar repositories for WhoSpeaks
Users that are interested in WhoSpeaks are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Real-time speaker diarization using straightforward, intuitive logic - High accuracy thanks to SpeechBrain/Pyannote-WeSpeaker models☆32May 7, 2026Updated 3 months ago
- Command Your World with Voice☆813Jun 17, 2025Updated last year
- Tr-VAD: An Efficient Transformer based Voice Activity Detection Model☆18Aug 1, 2024Updated 2 years ago
- Simulates talk with an AI that can express emotions☆89Apr 4, 2026Updated 4 months ago
- A python package to build AI-powered real-time audio applications☆2,013Jun 19, 2026Updated last month
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Transcribe desktop audio/computer audio in real-time and locally (Streaming ASR), using TorchAudio and Emformer-RNNT model for inference,…☆14May 7, 2024Updated 2 years ago
- Provide Gradio custom components to make the diarization-based audio labeling process easier and faster.☆72Apr 22, 2026Updated 3 months ago
- Speaker Diarization with Transformers☆70Jun 8, 2025Updated last year
- Private, real-time speaker diarization on hardware you control. See who is speaking as it happens, no third-party cloud required.☆18Jul 25, 2026Updated 2 weeks ago
- Very fast, accurate speaker diarization☆289Updated this week
- Roomey is a multi-purpose Voice Agent designed to run your personal and business life.☆67Jun 15, 2025Updated last year
- Low latency ai companion voice talk in 60 lines of code using faster_whisper and elevenlabs input streaming☆321Jun 17, 2025Updated last year
- DiariZen Explained: A Tutorial for the Open Source State-of-the-Art Speaker Diarization Pipeline.☆23Updated this week
- Real-time processing and delivery of sentences from a continuous stream of characters or text chunks.☆84Jul 16, 2026Updated 3 weeks ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Converts text to speech in realtime☆4,005Aug 2, 2026Updated last week
- Simple PyTorch Denoisers for Waveform Audio☆43Jul 28, 2026Updated last week
- Cog implementation of transcribing + diarization pipeline with Whisper & Pyannote☆237Jun 11, 2026Updated 2 months ago
- PAFTS : Library That Preprocessing Audio For TTS.☆27Nov 15, 2024Updated last year
- Blind Source Separation and Dereverberation☆21Mar 26, 2021Updated 5 years ago
- auto fine tune of models with synthetic data☆78Feb 14, 2024Updated 2 years ago
- This repo contains script to download MUSIC dataset from youtube☆12Jan 19, 2024Updated 2 years ago
- An application-layer router for Skupper networks☆21Updated this week
- Multi-Stage Face-Voice Association Learning with Keynote Speaker Diarization (ACM MM 2024)☆22Jul 25, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆14Aug 22, 2024Updated last year
- Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper☆5,616Feb 23, 2026Updated 5 months ago
- 🎹 pyannote + 🗒 notebook = pyannotebook☆27Jun 12, 2023Updated 3 years ago
- FastAPI WebSocket server for the OpenVoice text-to-speech model.☆12Jun 6, 2024Updated 2 years ago
- Identity verification from speech☆19Jul 19, 2022Updated 4 years ago
- [Colab Demo Code] OneFormer: One Transformer to Rule Universal Image Segmentation.☆14May 24, 2023Updated 3 years ago
- C++ version of pyannote audio speaker diarizaiton pipeline☆22Feb 14, 2024Updated 2 years ago
- Get started using Deepgram's Live Transcription with this Flask demo app☆46Jul 30, 2026Updated last week
- Authenticating proxy server for connecting to 3rd party APIs☆17Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Some comprehensive papers about speaker diarization☆369Mar 24, 2026Updated 4 months ago
- An unofficial implementation of the Personal VAD speaker-conditioned voice activity detection method. Bachelor's thesis project.☆90Sep 22, 2022Updated 3 years ago
- Speech Diarization for scrum automation☆111Jul 27, 2023Updated 3 years ago
- 💬 ASR FastAPI server using faster-whisper and Multi-Scale Auto-Tuning Spectral Clustering for diarization.☆219Oct 30, 2024Updated last year
- Make Qwen3 Think like Gemini 2.5 Pro | Open webui function☆25May 10, 2025Updated last year
- ☆13May 23, 2024Updated 2 years ago
- Official repository for the "Powerset multi-class cross entropy loss for neural speaker diarization" paper published in Interspeech 2023.☆96Oct 18, 2023Updated 2 years ago