A lightweight end-of-utterance detection model fine-tuned on SmolLM2-135M, optimized for Raspberry Pi and low-power devices.
☆65Mar 20, 2026Updated 6 months ago
Alternatives and similar repositories for turnsense
Users that are interested in turnsense are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Script to demonstrate how to use a Language Model for Semantic Turn Detection. Refer to blog post for full details.☆19May 9, 2025Updated last year
- Open-Source Turn-Taking Detection Model and Dataset for Full-Duplex Spoken Dialogue Systems☆145Jan 25, 2026Updated 8 months ago
- ☆21Mar 7, 2025Updated last year
- ☆1,599Jan 29, 2026Updated 7 months ago
- Vogent Turn: fast, open-source turn-detection for Voice AI applications☆53Oct 28, 2025Updated 11 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- High-performance, semantic turn detection for conversational AI☆46Oct 1, 2025Updated 11 months ago
- TEN VAD low-latency voice activity detection for real-time streaming, integrated with livekit-agents☆26Nov 13, 2025Updated 10 months ago
- Lightweight wrapper for Silero VAD using internal ONNX Runtime and with no python package dependencies☆17Nov 25, 2024Updated last year
- Voice activity engine benchmark framework☆23Jan 14, 2026Updated 8 months ago
- Source code for ASRU 2019 paper "Adapting Pretrained Transformer to Lattices for Spoken Language Understanding"☆10Jul 8, 2020Updated 6 years ago
- Turn detection for full-duplex dialogue communication☆616Dec 26, 2025Updated 9 months ago
- Multi-talker ASR based on DiCoW with Serialized Output Training☆21Sep 18, 2025Updated last year
- ☆122Jul 26, 2026Updated 2 months ago
- ☆29Feb 23, 2026Updated 7 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A SOTA Industrial-Grade Voice Activity Detection & Audio Event Detection, supporting 100+ languages, outperforming Silero-VAD, TEN-VAD, F…☆539May 6, 2026Updated 4 months ago
- Deep neural approach to Boundary and Disfluency Detection - Based on my Master's work☆20Jul 25, 2024Updated 2 years ago
- ☆42Mar 31, 2026Updated 5 months ago
- ICASSP2026 HumDial Challenge☆53May 28, 2026Updated 3 months ago
- ☆18Jun 25, 2026Updated 3 months ago
- ☆21Mar 6, 2026Updated 6 months ago
- ☆32Updated this week
- A Fully Self-Hosted Solution for Full-Duplex Voice Interaction☆589Sep 28, 2025Updated last year
- Transformer based ASR Engine.☆13Aug 23, 2021Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- GitHub Copilot Dev Day × 이화여자대학교 KING 게임 동아리 | 2025년 4월 13일 (월) 19:00~21:00 | Microsoft 서울 광화문 사옥☆25Apr 6, 2026Updated 5 months ago
- A Benchmark for Evaluating Turn-Taking and Overlap Handling in Full-Duplex Spoken Dialogue Models☆301May 20, 2026Updated 4 months ago
- Plug-and-play streaming semantic VAD for real-time full-duplex spoken dialogue systems.☆319Jul 17, 2026Updated 2 months ago
- ROS node and utilities for audio streams.☆13Jul 10, 2026Updated 2 months ago
- ☆19Jan 5, 2020Updated 6 years ago
- Attention-based model for keywords spotting☆19Aug 9, 2021Updated 5 years ago
- Keyword spotting for audio with attention (KWS model for audio)☆18Jul 15, 2021Updated 5 years ago
- ASLP Summer Inter@NPU☆13Jul 30, 2024Updated 2 years ago
- Conversational AI cookbook for developers — exploring real-time voice agents, streaming, and orchestration. 对话式 AI 开发者手册:探索实时语音、编排与工程实践。☆23Nov 13, 2025Updated 10 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- X-Talk is an open-source full-duplex cascaded spoken dialogue system framework enabling low-latency, interruptible, and human-like speech…☆245Updated this week
- fd-sds☆21Apr 8, 2026Updated 5 months ago
- ☆63Jun 2, 2026Updated 3 months ago
- ☆38May 19, 2026Updated 4 months ago
- Official PyTorch code for Deep Audio-Signal Holistic Embeddings☆206Nov 7, 2025Updated 10 months ago
- Eureka-Audio: A 1.7B lightweight audio–language model that matches 7B–30B models on ASR, audio understanding, and paralinguistic reasonin…☆46Apr 11, 2026Updated 5 months ago
- We propose C2SER, a novel audio-language model designed to enhance the stability and accuracy of speech emotion recognition through conte…☆49Mar 3, 2025Updated last year