Reading notes of speech or deep learning related papers, including Automatic Speech Recognition (ASR), Speech Enhancement and Dereverberation (SED), Speech Separation (SS), Sound Source Localization (SSL) and some other audio signal processing topics.
☆30Jun 8, 2023Updated 3 years ago
Alternatives and similar repositories for Paper-Reading-Notes
Users that are interested in Paper-Reading-Notes are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ResNet-STFT Model for Sound Source Localization☆20Aug 25, 2022Updated 4 years ago
- [WIP]Direction based Multi-Channel Speech Separation☆14Jan 25, 2024Updated 2 years ago
- An unofficial implementation of Lite-RTSE, a cost-effective lite model for real-time speech enhancement☆15Nov 19, 2023Updated 2 years ago
- We design a spectral compression mapping (SCM) for full-band speech enhancement, and propose a two-stage stream named MHA-DPCRN☆24Jul 4, 2022Updated 4 years ago
- Official PyTorch implementation of the Interspeech 2023 paper☆32Jul 5, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A MATLAB implementation of “Multiple Sound Source Counting and Localization Based on TF-Wise Spatial Spectrum Clustering” [TASLP 2019]☆11Oct 23, 2023Updated 2 years ago
- ☆16Sep 19, 2024Updated 2 years ago
- Nested U-Net with two-level skip connections for speech enhancement☆38Dec 18, 2023Updated 2 years ago
- A python implementation of “SRP-DNN: Learning Direct-Path Phase Difference for Multiple Moving Sound Source Localization” [ICASSP 2022]☆69Sep 28, 2024Updated 2 years ago
- ☆24Feb 28, 2023Updated 3 years ago
- Graph Neural Networks for Sound Source Localization☆30Oct 31, 2023Updated 2 years ago
- The implementation of TaylorBeamformer, which is in submission to Interspeech2022☆49Jun 10, 2022Updated 4 years ago
- A python implementation of “Self-Supervised Learning of Spatial Acoustic Representation with Cross-Channel Signal Reconstruction and Mult …☆41Oct 11, 2024Updated 2 years ago
- Speech Separation☆21Mar 7, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆11Aug 5, 2022Updated 4 years ago
- This is the official implementation of the LiSenNet☆165Nov 15, 2024Updated last year
- Pytorch implementation of DPCRN☆29Mar 31, 2024Updated 2 years ago
- ☆13Jun 24, 2021Updated 5 years ago
- ☆21Apr 27, 2024Updated 2 years ago
- This is the official implementation of PGUSE☆43Jun 7, 2025Updated last year
- Necessary and Sufficient Conditions for Observability of SLAM-based Microphone Array Calibration and Sound Source Localization☆14Mar 23, 2021Updated 5 years ago
- This is the microphone array generalization investigation based on previous Narrow Band Deep Filtering methods.☆38Mar 12, 2024Updated 2 years ago
- This is an unofficial Pytorch implementation of the DTLN model repository, which contains denoising and inference code for the DTLN model…☆24Jun 18, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- 把 wave-u-net 网络应用于语音增强领域中☆14May 29, 2020Updated 6 years ago
- Data simulation scripts for paper "Target Sound Extraction with Variable Cross-modality Clues"☆17May 19, 2023Updated 3 years ago
- Implementation of CGMM-MVDR beamforming used for Clarity challenge☆15Jan 14, 2022Updated 4 years ago
- ☆19Apr 1, 2020Updated 6 years ago
- Full implementation of "End-to-end microphone permutation and number invariant multi-channel speech separation" (Interspeech 2020)☆76Sep 14, 2021Updated 5 years ago
- ☆18Feb 1, 2026Updated 8 months ago
- Implementation of paper "DPCRN: Dual-Path Convolution Recurrent Network for Single Channel Speech Enhancement"☆237Apr 22, 2024Updated 2 years ago
- Communication-Cost Aware Microphone Selection For Neural Speech Enhancement with Ad-hoc Microphone Arrays☆17Nov 20, 2020Updated 5 years ago
- pre-process script for timit data for dnn-aec works☆38Mar 3, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Feedforward Sequential Memory Networks☆18Aug 2, 2022Updated 4 years ago
- The Official PyTorch Implementation of FN-SSL & IPDnet for Sound Source Localization [INTERSPEECH2023 & TASLP2024]☆166Mar 10, 2026Updated 7 months ago
- This is the official implementation of the SEMamba paper. (Accepted to IEEE SLT 2024)☆277Dec 12, 2025Updated 9 months ago
- ☆15Oct 12, 2023Updated 2 years ago
- bin2bin, a Time-Frequency Generative Adversarial based method for Audio Packet Loss Concealment☆17Dec 29, 2023Updated 2 years ago
- Baseline method for audio-visual sound event localization and detection task of DCASE 2023 challenge☆69Mar 19, 2025Updated last year
- ☆19Oct 26, 2023Updated 2 years ago