We implemented the DEMUCS model for speech enhancement in the time-frequency domain, and additionally implemented HD-DEMUCS.
☆34Nov 8, 2023Updated 2 years ago
Alternatives and similar repositories for DEMUCS-for-Speech-Enhancement
Users that are interested in DEMUCS-for-Speech-Enhancement are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- In this project, we will perform 12-lead ECG Multi-label Classification. Specifically, we will design a multi-model utilizing the charact…☆11Aug 26, 2024Updated last year
- Nested U-Net with two-level skip connections for speech enhancement☆38Dec 18, 2023Updated 2 years ago
- A Combined ResNet-DenseNet Architecture with ResU Blocks (ResU-Dense) for 12-lead ECG Abnormality Classification☆24Jul 16, 2024Updated 2 years ago
- 2D residual U-Net (ResUNet) and a lead combiner (LC) for 12-lead ECG Abnormality Classification☆15Jan 4, 2024Updated 2 years ago
- Causal Speech Enhancement Based on a Two-Branch Nested U-Net Architecture Using Self-Supervised Speech Embeddings☆21Jun 6, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Real-time speech enhancement mobile app using Nested U-Net☆55Oct 6, 2023Updated 2 years ago
- Real-Time Noise Reducer in Android☆35Dec 11, 2022Updated 3 years ago
- DNN-based SE in the frequency domain using Pytorch. You can test some state-of-the-art networks using T-F masking or spectral mapping met…☆61Apr 2, 2022Updated 4 years ago
- 6 DoF Directional Room Impulse Response (RIR) with Dense Loudspeaker Grid☆17Aug 31, 2023Updated 2 years ago
- PyTorch-based room impulse response (RIR) simulation toolkit with dynamic scenes, GPU acceleration.☆23Updated this week
- Speech enhancement by time-varying pitch-dependent filtering of harmonics☆27Jul 3, 2014Updated 12 years ago
- Generator for anechoic, non-stationary noise signals☆12Aug 12, 2022Updated 3 years ago
- Spherical residual vector quantization (SRVQ)☆31Aug 25, 2024Updated last year
- DCCRN with various loss functions☆103Sep 29, 2022Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Room impulse response simulation for various array architectures using Monte-Carlo simulation and quaternions (Python)☆18Feb 25, 2026Updated 5 months ago
- Official implementation of "Wave-Trainer-Fit: Neural Vocoder with Trainable Prior and Fixed-Point Iteration towards High-Quality Speech G…☆16Feb 6, 2026Updated 5 months ago
- FINALLY: Fast and universal speech enhancement model delivering studio-quality audio for a wide range of recordings.☆28Apr 1, 2026Updated 3 months ago
- Lightweight Korean TTS Model based on FastSpeech2☆15Mar 4, 2026Updated 4 months ago
- This is the code for the paper "On Interference-Rejection using Riemannian Geometry for Direction of Arrival Estimation", A. Bar and R. T…☆21Oct 23, 2023Updated 2 years ago
- Official repository for Mamba-based Segmentation Model for Speaker Diarization☆47May 13, 2025Updated last year
- Generating non-stationary multi-sensor signals under a spatial coherence constraint (MATLAB)☆50Sep 25, 2024Updated last year
- Official PyTorch implementation of "RVAE-EM: Generative speech dereverberation based on recurrent variational auto-encoder and convolutiv…☆51Mar 6, 2025Updated last year
- Generating non-stationary multi-sensor signals under a spatial coherence constraint (Python)☆31Apr 12, 2026Updated 3 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Model configurations for scaling SE models in the paper "Beyond Performance Plateaus: A Comprehensive Study on Scalability in Speech Enha…☆41Aug 7, 2024Updated last year
- Unofficial Implementation of "Liu, W., Li, A., Wang, X., Yuan, M., Chen, Y., Zheng, C., & Li, X. (2022). A Neural Beamspace-Domain Filter…☆19Oct 21, 2022Updated 3 years ago
- SS2 HRTF Dataset - Reality Labs Research Audio☆18May 22, 2026Updated 2 months ago
- [WIP]Direction based Multi-Channel Speech Separation☆14Jan 25, 2024Updated 2 years ago
- ☆14Nov 28, 2022Updated 3 years ago
- A description of "RealMAN: A Real-Recorded and Annotated Microphone Array Dataset for Dynamic Speech Enhancement and Localization" [NeurI…☆175Apr 29, 2025Updated last year
- ASLP Summer Inter@NPU☆13Jul 30, 2024Updated last year
- ☆17Mar 30, 2023Updated 3 years ago
- ☆21Jul 16, 2023Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆19Oct 9, 2025Updated 9 months ago
- a lightweight network for monaural speech enhancement☆58Oct 12, 2023Updated 2 years ago
- Single channel speech source separation by diffusion process (ICASSP 2023)☆126Mar 15, 2024Updated 2 years ago
- Higher-Order Ambisonics Codec for Spatial Audio☆45May 11, 2025Updated last year
- An official documentation of the paper <Wave-U-Mamba: An End-To-End Framework For High-Quality And Efficient Speech Super Resolution>.☆26Oct 29, 2025Updated 9 months ago
- ☆17Sep 12, 2023Updated 2 years ago
- Code for the paper "DSpAST: Disentangled Representations for Spatial Audio Reasoning with Large Language Models"☆17Oct 23, 2025Updated 9 months ago