☆19Jul 12, 2020Updated 6 years ago
Alternatives and similar repositories for VCTK-2Mix
Users that are interested in VCTK-2Mix are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An open source dataset for source separation☆499Feb 9, 2024Updated 2 years ago
- ☆73Feb 15, 2021Updated 5 years ago
- ☆15Sep 6, 2021Updated 4 years ago
- This repo contains conv-tasnet for basis-melgan. If you want to get code of basis-melgan, please refer to FastVocoder.☆21Jul 21, 2021Updated 5 years ago
- ☆25Sep 30, 2019Updated 6 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Permutation invariant training in PyTorch☆13Oct 2, 2020Updated 5 years ago
- Libri-CSS: dataset and evaluation pipeline☆157Jan 18, 2023Updated 3 years ago
- the MEX wrapper for PESQ (Perceptual Evaluation of Speech Quality)☆15May 10, 2019Updated 7 years ago
- This repo contains required files for the INTERSPEECH 2022 Audio Deep Packet Loss Concealment (PLC) Challenge.☆92Feb 13, 2026Updated 5 months ago
- ☆53May 15, 2025Updated last year
- Source code and audio samples for the paper "MixCycle: Unsupervised Speech Separation via Cyclic Mixture Permutation Invariant Training"☆24Jun 21, 2026Updated last month
- PyTorch implementations of neural network models for keyword spotting☆11Oct 19, 2020Updated 5 years ago
- Perceptual Contrast Stretching on Target Feature for Speech Enhancement (Accepted by INTERSPEECH 2022)☆73May 11, 2024Updated 2 years ago
- Performance-oriented implementation of independent vector analysis for blind source separation.☆26Mar 26, 2020Updated 6 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆25Nov 23, 2021Updated 4 years ago
- ☆25Jul 7, 2022Updated 4 years ago
- ☆20Nov 22, 2020Updated 5 years ago
- Functions for creating speech features in MATLAB.☆14Jul 7, 2020Updated 6 years ago
- Zero-Shot Blind Audio Bandwidth Extension☆27May 25, 2023Updated 3 years ago
- A Temporal-Spectral Generative Adversarial Network based End-to-end Packet Loss Concealment for Wideband Speech Transmission☆32Apr 27, 2022Updated 4 years ago
- The audio demos with respect to the paper "DBT-Net: Dual-branch federative magnitude and phase estimation with attention-in-attention tra…☆30Jul 25, 2022Updated 3 years ago
- This repo contains some object detection algorithms and techniques (Not ML algorithms). This is aimed to get coordinates, width, height, …☆12Nov 26, 2020Updated 5 years ago
- ☆93Jun 9, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆23May 15, 2023Updated 3 years ago
- Unofficial Multi-microphone complex spectral mapping for utterance-wise and continuous speech separation(MISO-BF-MISO)☆52Jan 13, 2022Updated 4 years ago
- A simple package for Guided source separation (GSS)☆134May 20, 2024Updated 2 years ago
- ☆24Apr 25, 2022Updated 4 years ago
- target speaker extraction and verification for multi-talker speech☆210Jan 24, 2021Updated 5 years ago
- PyTorch implementation of the ICASSP-24 paper: "Improving Audio Captioning Models with Fine-grained Audio Features, Text Embedding Superv…☆41Jan 6, 2024Updated 2 years ago
- The implementation of "A Recursive Network with Dynamic Attention for Monaural Speech Enhancement"☆80Dec 8, 2022Updated 3 years ago
- The updated version of TDAA model.☆14Jul 2, 2020Updated 6 years ago
- Model configurations for scaling SE models in the paper "Beyond Performance Plateaus: A Comprehensive Study on Scalability in Speech Enha…☆41Aug 7, 2024Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Complex-valued Spatial Autoencoders for Multichannel Speech Enhancement☆33Apr 15, 2022Updated 4 years ago
- The implementation of "Dual-branch Attention-In-Attention Transformer for single-channel speech enhancement"☆126Jun 29, 2022Updated 4 years ago
- This is a single-speaker neural text-to-speech (TTS) system capable of training in a end-to-end fashion. It is inspired by the Tacotron a…☆12Dec 28, 2018Updated 7 years ago
- Psychoacoustic Calibration for Efficient Neural Audio Coding☆26Sep 26, 2023Updated 2 years ago
- ☆146Oct 25, 2021Updated 4 years ago
- Implementation of F5-TTS in MLX☆14Dec 13, 2024Updated last year
- wsj0-{2, 3, 4, 5} mix generation scripts, in Python.☆79Mar 17, 2021Updated 5 years ago