☆21Jul 15, 2024Updated 2 years ago
Alternatives and similar repositories for codecformer
Users that are interested in codecformer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation of Efficient Speech Separation Framework Based on Neural State-Space Models☆28Feb 25, 2026Updated 5 months ago
- offical code for Dense-TSNet☆12Sep 17, 2024Updated last year
- Power-Guided Grouped SRU for Real-Time Causal Audio-Visual Speech Separation☆30Jul 20, 2026Updated 3 weeks ago
- ☆12Sep 12, 2024Updated last year
- ☆53Sep 10, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Source code and demo for INTERSPEECH 2024 paper: Noise-robust Speech Separation with Fast Generative Correction☆52Nov 19, 2024Updated last year
- ☆69Aug 16, 2023Updated 2 years ago
- Microphone Array Real-time System☆13Jun 7, 2017Updated 9 years ago
- Aty-TTS: Improving fairness for spoken language understanding in atypical speech with Text-to-Speech☆12May 14, 2025Updated last year
- ☆13Dec 7, 2022Updated 3 years ago
- Official Implementation of TSELM: Target speaker extraction using discrete tokens and language models☆61Apr 14, 2025Updated last year
- The source code of Tim-TSENet☆15Apr 22, 2022Updated 4 years ago
- Train no-reference speech quality estimators with multiple datasets via learned, per-dataset alignments.☆18Aug 1, 2025Updated last year
- A toolkit for researchers in the multimodal sound separation.☆16Oct 20, 2023Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆32Jan 9, 2024Updated 2 years ago
- unofficial implementation of "CPTNN: CROSS-PARALLEL TRANSFORMER NEURAL NETWORK FOR TIME-DOMAIN SPEECH ENHANCEMENT"☆15Nov 14, 2023Updated 2 years ago
- Code for paper "Noise-aware Speech Enhancement using Diffusion Probabilistic Model"☆89Jun 10, 2024Updated 2 years ago
- ☆226Dec 5, 2024Updated last year
- ☆37Jan 6, 2026Updated 7 months ago
- Implementation of the paper "Variable Bitrate Residual Vector Quantization for Audio Coding"☆11Apr 10, 2025Updated last year
- Whisper Speech Quality Assessment (WhiSQA)☆16Apr 14, 2026Updated 3 months ago
- ☆10Mar 22, 2023Updated 3 years ago
- Streaming Vocos☆31Jun 10, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official source code of the INTERSPEECH 2023 paper: "Audio-Visual Speech Separation in Noisy Environments with a Lightweight Iterative Mo…☆20Sep 1, 2023Updated 2 years ago
- Automatic speech annotator processing speech with voice activaty detection, overlapping speech detection, speaker diarization and automat…☆33Jun 14, 2024Updated 2 years ago
- Evaluation tool used in the BigVSAN paper☆14Mar 22, 2024Updated 2 years ago
- Unofficial implementation of wavenext vocoder☆59Aug 28, 2024Updated last year
- a Neural Vocoder supporting Ring Attention, Conformer and NSF.☆25Aug 1, 2025Updated last year
- ☆23Jul 18, 2026Updated 3 weeks ago
- Official code for MUSE: Flexible Voiceprint Receptive Fields and Multi-Path Fusion Enhanced Taylor Transformer for U-Net-based Speech Enh…☆58Mar 5, 2025Updated last year
- Transformer with Local Modeling by Convolution for Speech Separation and Enhancement☆135Aug 8, 2025Updated last year
- ☆45Sep 19, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Papez: Resource-Efficient Speech Separation with Auditory Working Memory (ICASSP 2023)☆22Jun 25, 2023Updated 3 years ago
- This is the official implementation of our multi-channel multi-speaker multi-spatial neural audio codec architecture.☆55Mar 17, 2025Updated last year
- A neural speech codec based on discrete WavLM representations☆26Aug 28, 2024Updated last year
- Pytorch implementation of the paper : A Global-local Attention Framework for Weakly Labelled Audio Tagging.☆13Feb 6, 2021Updated 5 years ago
- Code for the "NoiseBandNet: Controllable Time-Varying Neural Synthesis of Sound Effects Using Filterbanks" paper.☆39Jul 8, 2024Updated 2 years ago
- ☆59Apr 24, 2024Updated 2 years ago
- Code for paper Learning Audio-Visual Dereverberation☆32Aug 10, 2022Updated 4 years ago