Implementation of "Personal VAD 2.0: Optimizing Personal Voice Activity Detection for On-Device Speech Recognition"
☆15Jun 9, 2026Updated last month
Alternatives and similar repositories for Personal-vad-2.0
Users that are interested in Personal-vad-2.0 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A repository for code used to produce the results the ICASSP 2024 paper: "SELF-SUPERVISED PRETRAINING FOR ROBUST PERSONALIZED VOICE ACTIV…☆25Nov 25, 2024Updated last year
- FINALLY: Fast and universal speech enhancement model delivering studio-quality audio for a wide range of recordings.☆28Apr 1, 2026Updated 3 months ago
- An unofficial implementation of the Personal VAD speaker-conditioned voice activity detection method. Bachelor's thesis project.☆89Sep 22, 2022Updated 3 years ago
- Attention-Based Encoder-Decoder Target-Speaker Voice Activity Detection for Robust Speaker Diarization☆31Sep 22, 2025Updated 9 months ago
- FireRedChat pVAD plugin for LiveKit Agents☆22Sep 16, 2025Updated 10 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- The code about “LABNet: A Lightweight Attentive Beamforming Network for Ad-hoc Multichannel Microphone Invariant Real-Time Speech Enhance…☆49Oct 10, 2025Updated 9 months ago
- ☆70Jul 5, 2025Updated last year
- ☆26Aug 29, 2025Updated 10 months ago
- ASLP Summer Inter@NPU☆12Jul 30, 2024Updated last year
- 音频处理小工具☆14Jun 4, 2026Updated last month
- This is the official implementation of PGUSE☆41Jun 7, 2025Updated last year
- Spherical residual vector quantization (SRVQ)☆31Aug 25, 2024Updated last year
- ☆29Jan 22, 2026Updated 5 months ago
- ☆26Mar 31, 2026Updated 3 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- PyTorch-based room impulse response (RIR) simulation toolkit with dynamic scenes, GPU acceleration.☆22Feb 18, 2026Updated 5 months ago
- ☆22Oct 17, 2024Updated last year
- ☆62Apr 11, 2022Updated 4 years ago
- Ablation study of local spectral attention (LSA) for full-band speech enhancement (SE)☆28Sep 16, 2023Updated 2 years ago
- pre-process script for timit data for dnn-aec works☆38Mar 3, 2022Updated 4 years ago
- Subband kalman filter for echo cancellation☆67Jul 6, 2023Updated 3 years ago
- A pytoch lightning training implementation of SLAM-ASR☆11Nov 17, 2025Updated 8 months ago
- Generate synthetic wind noise signals based on a wind speed profile (Python)☆52Apr 23, 2024Updated 2 years ago
- Speed-optimized streaming neural speech enhancement network☆133Jul 3, 2026Updated 2 weeks ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- This is an evolving repo for the paper “From Turn-Taking to Synchronous Dialogue: A Survey of Full-Duplex Spoken Language Models ”A compr…☆25Dec 23, 2025Updated 6 months ago
- ☆24Jul 10, 2025Updated last year
- ☆26Jun 10, 2026Updated last month
- TS-SEP: Joint Diarization and Separation Conditioned on Estimated Speaker Embeddings☆43Oct 27, 2025Updated 8 months ago
- Implementation of CGMM-MVDR beamforming used for Clarity challenge☆14Jan 14, 2022Updated 4 years ago
- A demo-level low-latency, high-throughput inference engine for whisper☆20Nov 9, 2025Updated 8 months ago
- ☆16Mar 19, 2026Updated 4 months ago
- Codebase of the submitted work in ICASSP 2023☆14Nov 30, 2022Updated 3 years ago
- 在Android上运行人脸表情识别的tflite模型☆12Apr 7, 2021Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- PASE: Phonologically Anchored Speech Enhancer☆84Updated this week
- Python的音频工具☆16Dec 5, 2025Updated 7 months ago
- ☆34Nov 29, 2022Updated 3 years ago
- AEC Challenge☆14Nov 12, 2021Updated 4 years ago
- Official Repository for "Efficient Vocal Source Separation Through Windowed RoFormer"☆45Oct 30, 2025Updated 8 months ago
- A list of papers for child ASR☆54Oct 8, 2024Updated last year
- FastWave is a lightweight diffusion model for general audio super-resolution (any -> 48 kHz). SOTA quality reconstruction metrics with ju…☆18May 16, 2026Updated 2 months ago