☆25Oct 4, 2022Updated 3 years ago
Alternatives and similar repositories for vctk-silence-labels
Users that are interested in vctk-silence-labels are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- maum-ai.github.io☆15Jun 12, 2026Updated last month
- Towards High-Quality and Efficient Speech Bandwidth Extension with Parallel Amplitude and Phase Prediction☆13Jul 22, 2024Updated 2 years ago
- NU-Wave: A Diffusion Probabilistic Model for Neural Audio Upsampling☆37May 25, 2021Updated 5 years ago
- SANE-TTS: Stable And Natural End-to-End Multilingual Text-to-Speech☆11Jun 30, 2023Updated 3 years ago
- ☆20Jul 13, 2022Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Pitch-shift audio clips quickly with PyTorch (CUDA supported)! Additional utilities for searching efficient transformations are included.☆139Sep 25, 2024Updated last year
- Code for the CVPR2021 workshop paper "Noise Conditional Flow Model for Learning the Super-Resolution Space"☆64Jun 21, 2021Updated 5 years ago
- A repository for benchmarking neural vocoders by their quality and speed.☆213May 30, 2025Updated last year
- Official repository for the paper "Chunked Autoregressive GAN for Conditional Waveform Synthesis"☆193Dec 8, 2022Updated 3 years ago
- ☆171Jul 25, 2022Updated 4 years ago
- PyTorch Implementation of Google Brain's WaveGrad 2: Iterative Refinement for Text-to-Speech Synthesis☆68Aug 3, 2021Updated 4 years ago
- ☆48Updated this week
- Evaluation and Benchmarking of Speech Super-resolution Methods☆157Jun 17, 2022Updated 4 years ago
- NU-Wave 2: A General Neural Audio Upsampling Model for Various Sampling Rates [WIP]☆25Jul 5, 2022Updated 4 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Simple sinc interpolation in PyTorch.☆15Jul 8, 2023Updated 3 years ago
- UnivNet: A Neural Vocoder with Multi-Resolution Spectrogram Discriminators for High-Fidelity Waveform Generation☆76Aug 30, 2021Updated 4 years ago
- A solution to denoising and separating for two-speaker-mixed noisy speech, using a BSRNN inspired network.☆15Aug 22, 2023Updated 2 years ago
- The official implementation of VAENAR-TTS, a VAE based non-autoregressive TTS model.☆144Jul 8, 2021Updated 5 years ago
- The official implementation of the Interspeech 2021 paper WSRGlow: A Glow-based Waveform Generative Model for Audio Super-Resolution.☆127Sep 7, 2021Updated 4 years ago
- ☆87May 21, 2023Updated 3 years ago
- ☆18Nov 10, 2019Updated 6 years ago
- Based on https://github.com/fatchord/WaveRNN☆24May 3, 2020Updated 6 years ago
- Official implementation of "Avocodo: Generative Adversarial Network for Artifact-Free Vocoder" (AAAI2023)☆154Feb 1, 2023Updated 3 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- 🤫A Lightweight One-Shot Whisper to Normal Voice Conversion Model Using Distillation of Self-Supervised Features☆25Dec 10, 2025Updated 7 months ago
- ICASSP2022 TTS&VC Summary☆13Jun 9, 2022Updated 4 years ago
- Official Code for Assem-VC @ICASSP2022☆269May 16, 2022Updated 4 years ago
- Versatile Evaluation of Speech and Audio☆425Jul 21, 2026Updated last week
- ICASSP 2023 Accepted☆191May 6, 2024Updated 2 years ago
- Info for prospective PhD students for Chris Donahue's lab at CMU starting Fall 23.☆12Nov 13, 2022Updated 3 years ago
- Implementation of CGMM-MVDR beamforming used for Clarity challenge☆14Jan 14, 2022Updated 4 years ago
- NU-Wave: A Diffusion Probabilistic Model for Neural Audio Upsampling @ INTERSPEECH 2021☆283Jul 22, 2022Updated 4 years ago
- A python wrapper for REAPER☆81Jan 22, 2025Updated last year
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- ☆12Nov 7, 2024Updated last year
- An official reimplementation of the method described in the INTERSPEECH 2021 paper - Speech Resynthesis from Discrete Disentangled Self-S…☆416Aug 29, 2023Updated 2 years ago
- Official data preparation scripts for the URGENT 2024 Challenge☆90May 21, 2025Updated last year
- Audio samples for the paper 'Phase-aware music super-resolution using generative adversarial networks'☆14May 15, 2020Updated 6 years ago
- NU-Wave 2: A General Neural Audio Upsampling Model for Various Sampling Rates @ INTERSPEECH 2022☆312Sep 16, 2023Updated 2 years ago
- Labels for kiritan_singing data with extra resources for DNN-based singing voice synthesis (SVS) systems.☆28Dec 31, 2023Updated 2 years ago
- This is the main repository of open-sourced speech technology by Huawei Noah's Ark Lab.☆604Sep 18, 2023Updated 2 years ago