Real-time Speech Separation, Noise Suppression & Speaker Recognition
☆17Apr 17, 2019Updated 7 years ago
Alternatives and similar repositories for audiovision
Users that are interested in audiovision are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- PyTorch implementation of TinyWASE described in our paper "Compressing Speaker Extraction Model with Ultra-low Precision Quantization and…☆11Jun 28, 2021Updated 5 years ago
- ☆39Feb 23, 2022Updated 4 years ago
- ☆16Jun 15, 2022Updated 4 years ago
- Generalized RNN beamformer for speech separation☆19Jan 11, 2022Updated 4 years ago
- DCCRN: Deep Complex Convolution Recurrent Network☆15Nov 26, 2021Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- PyTorch implementation of WASE described in our ICASSP 2021: "Wase: Learning When to Attend for Speaker Extraction in Cocktail Party Envi…☆27Jan 11, 2022Updated 4 years ago
- ☆16Jan 20, 2021Updated 5 years ago
- PyTorch implementation of LiMuSE☆33Oct 11, 2022Updated 3 years ago
- microphone array speech generator (MASG) in room acoustic☆39Jan 2, 2020Updated 6 years ago
- ☆149Oct 25, 2021Updated 4 years ago
- Source code and audio demos for the paper "Audio Source Separation Using Variational Autoencoders and Weak Class Supervision"☆12Jun 21, 2026Updated 2 months ago
- Speech enhancement system for the CHiME-5 dinner party scenario☆111Feb 6, 2025Updated last year
- COVID-19: Face Mask Detector with OpenCV, Keras/TensorFlow, and Deep Learning☆12Dec 23, 2020Updated 5 years ago
- ☆38Jul 20, 2020Updated 6 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Classification of environmental sounds using first order statistics and GLCM (Gray-Level Co-Occurrence Matrix ) features of a spectrogram…☆25Jul 14, 2020Updated 6 years ago
- multi-scale time domain speaker extraction☆82Jun 7, 2021Updated 5 years ago
- A CNN-based audio denoiser☆10May 2, 2021Updated 5 years ago
- ☆17Nov 17, 2020Updated 5 years ago
- Implementation of "SpEx: Multi-Scale Time Domain Speaker Extraction Network".☆37Jul 19, 2020Updated 6 years ago
- A PyTorch implementation: "LASAFT-Net-v2: Listen, Attend and Separate by Attentively aggregating Frequency Transformation"☆33Apr 11, 2022Updated 4 years ago
- ☆53Jun 14, 2022Updated 4 years ago
- This repository is webrtc agc module demo.☆12Jan 23, 2019Updated 7 years ago
- ☆17Sep 12, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official PyTorch implementation of MVAE for audio source separation☆43Dec 21, 2022Updated 3 years ago
- Adaptive and Focusing Neural Layers for Multi-Speaker Separation Problem☆50Jul 7, 2018Updated 8 years ago
- C++ version of pyannote audio overlapped speech detection pipeline☆13Feb 14, 2024Updated 2 years ago
- ☆46Dec 5, 2019Updated 6 years ago
- bktree data structure with a Python interface for a CPP implementation☆13Jan 11, 2017Updated 9 years ago
- A pure-Python, bring-your-own-I/O implementation of HTTP/1.1☆14Oct 30, 2018Updated 7 years ago
- PyTorch based speaker embedding model☆16Apr 13, 2024Updated 2 years ago
- ☆15Sep 16, 2024Updated last year
- A c++ wrapper for the LAME library that reduces conversion of PCM (*.wav) to mp3 and vice versa to just two lines of codes.☆12Jan 8, 2015Updated 11 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- 语音算法相关资源汇总 Resource for Speech Processing || NEWS: official link of VoxCeleb fails recently and an external link is added for download☆60Jul 24, 2022Updated 4 years ago
- Python package for noise supression in audio based on DNN☆22Mar 24, 2023Updated 3 years ago
- ☆13Mar 22, 2021Updated 5 years ago
- The implementation of "Optimizing Shoulder to Shoulder: A Coordinated Sub-Band Fusion Model for Real-Time Full-Band Speech Enhancement"☆54Feb 16, 2023Updated 3 years ago
- An open-source speech separation and enhancement library☆214May 13, 2020Updated 6 years ago
- 💻 CMake function that wrap macdeployqt, deploy dmg and pkg.☆11Jan 8, 2026Updated 7 months ago
- offical code for Dense-TSNet☆12Sep 17, 2024Updated last year