🤫A Lightweight One-Shot Whisper to Normal Voice Conversion Model Using Distillation of Self-Supervised Features
☆25Dec 10, 2025Updated 7 months ago
Alternatives and similar repositories for distillw2n
Users that are interested in distillw2n are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A neural speech codec based on discrete WavLM representations☆26Aug 28, 2024Updated last year
- ☆17Apr 9, 2026Updated 3 months ago
- Official repository of UniPASE, a SOTA USE model☆51Jul 21, 2026Updated last week
- ☆36Dec 25, 2023Updated 2 years ago
- FNSE-SBGAN: Far-field Speech Enhancement with Schrödinger Bridge and Generative Adversarial Networks☆20May 12, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Implementation of "Improving Whispered Speech Recognition Performance using Pseudo-whispered based Data Augmentation"☆14Oct 31, 2024Updated last year
- Attention-Based Encoder-Decoder Target-Speaker Voice Activity Detection for Robust Speaker Diarization☆31Sep 22, 2025Updated 10 months ago
- Source code and audio samples for AFC-SPEX, an algorithm that can jointly perform acoustic feedback cancellation and speaker extraction.☆40Nov 7, 2025Updated 8 months ago
- PASE: Phonologically Anchored Speech Enhancer☆86Jul 15, 2026Updated last week
- ☆46Jan 14, 2025Updated last year
- ☆52Sep 10, 2024Updated last year
- A Lightweight Hybrid Dual Channel Speech Enhancement System under Low-SNR Conditions (Interspeech 2025)☆111Mar 13, 2026Updated 4 months ago
- ☆21Jul 18, 2026Updated last week
- Model configurations for scaling SE models in the paper "Beyond Performance Plateaus: A Comprehensive Study on Scalability in Speech Enha…☆41Aug 7, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Official code of SenSE.☆90Oct 30, 2025Updated 8 months ago
- Multi-Scale Temporal Frequency Convolutional Network With Axial Attention for Speech Enhancement☆233Sep 30, 2022Updated 3 years ago
- ☆59Apr 24, 2024Updated 2 years ago
- Official page of "DeFTAN-II: Efficient multichannel speech enhancement with subgroup processing", IEEE/ACM Transactions on Audio, Speech,…☆34Nov 21, 2024Updated last year
- This is the code and dataset repo for Interspeech 2024 paper "Target conversation extraction: Source separation using turn-taking dynamic…☆58Aug 15, 2025Updated 11 months ago
- Implementation of Sheffield entry for Clarity enhancement challenge.☆18Apr 19, 2022Updated 4 years ago
- ☆12Nov 7, 2024Updated last year
- semantic tokenizer for speech and music☆20Jul 6, 2025Updated last year
- Personalized AEC☆19Nov 3, 2022Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆70Jul 5, 2025Updated last year
- ☆39Feb 23, 2026Updated 5 months ago
- ☆21Jul 15, 2024Updated 2 years ago
- Pytorch implementation of DPCRN☆29Mar 31, 2024Updated 2 years ago
- Official Repository For VoxBlink2☆88Aug 13, 2024Updated last year
- Speech Separation☆21Mar 7, 2024Updated 2 years ago
- Fast algorithm for determined blind source separation with update of demixing filters with joint adjustment of the remaining sources.☆36Mar 22, 2021Updated 5 years ago
- This is the official implementation of the LiSenNet☆163Nov 15, 2024Updated last year
- Fully Quantized Neural Networks For Speech Enhancement☆65Feb 15, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Meanflow and multilingual for F5-TTS model☆16Aug 23, 2025Updated 11 months ago
- A repository for code used to produce the results the ICASSP 2024 paper: "SELF-SUPERVISED PRETRAINING FOR ROBUST PERSONALIZED VOICE ACTIV…☆25Nov 25, 2024Updated last year
- The Official PyTorch Implementation of "Mel-McNet: A Mel-Scale Framework for Online Multichannel Speech Enhancement" [Interspeech 2025]☆26May 14, 2026Updated 2 months ago
- An unofficial implementation of DeepVQE proposed by Microsoft Corp.☆148Mar 24, 2025Updated last year
- 一款隐身于 Mac 摄像头下方的智能提词器:专为视频录制、直播与会议设计,帮您保持自然眼神交流。支持苹果自带语音识别与本地 AI 大模型,能随着您的真实语速自动跟踪和滚动文案,彻底告别忘词与手动滑屏的烦恼。☆19Feb 24, 2026Updated 5 months ago
- ☆138Apr 24, 2023Updated 3 years ago
- Official repository for the paper "Exploring Resolution-Wise Shared Attention in Hybrid Mamba-U-Nets for Improved Cross-Corpus Speech Enh…☆19May 5, 2026Updated 2 months ago