We implemented the DEMUCS model for speech enhancement in the time-frequency domain, and additionally implemented HD-DEMUCS.
☆34Nov 8, 2023Updated 2 years ago
Alternatives and similar repositories for DEMUCS-for-Speech-Enhancement
Users that are interested in DEMUCS-for-Speech-Enhancement are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- In this project, we will perform 12-lead ECG Multi-label Classification. Specifically, we will design a multi-model utilizing the charact…☆12Aug 26, 2024Updated 2 years ago
- Nested U-Net with two-level skip connections for speech enhancement☆38Dec 18, 2023Updated 2 years ago
- A Combined ResNet-DenseNet Architecture with ResU Blocks (ResU-Dense) for 12-lead ECG Abnormality Classification☆24Jul 16, 2024Updated 2 years ago
- 2D residual U-Net (ResUNet) and a lead combiner (LC) for 12-lead ECG Abnormality Classification☆15Jan 4, 2024Updated 2 years ago
- Causal Speech Enhancement Based on a Two-Branch Nested U-Net Architecture Using Self-Supervised Speech Embeddings☆21Jun 6, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Real-time speech enhancement mobile app using Nested U-Net☆55Oct 6, 2023Updated 2 years ago
- Real-Time Noise Reducer in Android☆35Dec 11, 2022Updated 3 years ago
- DNN-based SE in the frequency domain using Pytorch. You can test some state-of-the-art networks using T-F masking or spectral mapping met…☆61Apr 2, 2022Updated 4 years ago
- 6 DoF Directional Room Impulse Response (RIR) with Dense Loudspeaker Grid☆17Aug 31, 2023Updated 3 years ago
- PyTorch-based room impulse response (RIR) simulation toolkit with dynamic scenes, GPU acceleration.☆23Jul 30, 2026Updated last month
- Speech enhancement by time-varying pitch-dependent filtering of harmonics☆28Jul 3, 2014Updated 12 years ago
- Generator for anechoic, non-stationary noise signals☆12Aug 12, 2022Updated 4 years ago
- Spherical residual vector quantization (SRVQ)☆31Aug 25, 2024Updated 2 years ago
- DCCRN with various loss functions☆103Sep 29, 2022Updated 3 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Room impulse response simulation for various array architectures using Monte-Carlo simulation and quaternions (Python)☆18Feb 25, 2026Updated 6 months ago
- Official implementation of "Wave-Trainer-Fit: Neural Vocoder with Trainable Prior and Fixed-Point Iteration towards High-Quality Speech G…☆16Feb 6, 2026Updated 7 months ago
- FINALLY: Fast and universal speech enhancement model delivering studio-quality audio for a wide range of recordings.☆29Apr 1, 2026Updated 5 months ago
- Lightweight Korean TTS Model based on FastSpeech2☆15Mar 4, 2026Updated 6 months ago
- This is the code for the paper "On Interference-Rejection using Riemannian Geometry for Direction of Arrival Estimation", A. Bar and R. T…☆21Oct 23, 2023Updated 2 years ago
- Official repository for Mamba-based Segmentation Model for Speaker Diarization☆47May 13, 2025Updated last year
- Generating non-stationary multi-sensor signals under a spatial coherence constraint (MATLAB)☆51Sep 25, 2024Updated last year
- Official PyTorch implementation of "RVAE-EM: Generative speech dereverberation based on recurrent variational auto-encoder and convolutiv…☆51Mar 6, 2025Updated last year
- Generating non-stationary multi-sensor signals under a spatial coherence constraint (Python)☆33Apr 12, 2026Updated 4 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Model configurations for scaling SE models in the paper "Beyond Performance Plateaus: A Comprehensive Study on Scalability in Speech Enha…☆42Aug 7, 2024Updated 2 years ago
- Unofficial Implementation of "Liu, W., Li, A., Wang, X., Yuan, M., Chen, Y., Zheng, C., & Li, X. (2022). A Neural Beamspace-Domain Filter…☆19Oct 21, 2022Updated 3 years ago
- SS2 HRTF Dataset - Reality Labs Research Audio☆18May 22, 2026Updated 3 months ago
- [WIP]Direction based Multi-Channel Speech Separation☆14Jan 25, 2024Updated 2 years ago
- ☆14Nov 28, 2022Updated 3 years ago
- A description of "RealMAN: A Real-Recorded and Annotated Microphone Array Dataset for Dynamic Speech Enhancement and Localization" [NeurI…☆178Apr 29, 2025Updated last year
- ASLP Summer Inter@NPU☆13Jul 30, 2024Updated 2 years ago
- ☆17Mar 30, 2023Updated 3 years ago
- ☆21Jul 16, 2023Updated 3 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆19Oct 9, 2025Updated 10 months ago
- a lightweight network for monaural speech enhancement☆59Oct 12, 2023Updated 2 years ago
- Single channel speech source separation by diffusion process (ICASSP 2023)☆127Mar 15, 2024Updated 2 years ago
- Higher-Order Ambisonics Codec for Spatial Audio☆45May 11, 2025Updated last year
- An official documentation of the paper <Wave-U-Mamba: An End-To-End Framework For High-Quality And Efficient Speech Super Resolution>.☆26Oct 29, 2025Updated 10 months ago
- ☆17Sep 12, 2023Updated 2 years ago
- Code for the paper "DSpAST: Disentangled Representations for Spatial Audio Reasoning with Large Language Models"☆17Oct 23, 2025Updated 10 months ago