☆21Jul 15, 2024Updated 2 years ago
Alternatives and similar repositories for codecformer
Users that are interested in codecformer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation of Efficient Speech Separation Framework Based on Neural State-Space Models☆29Feb 25, 2026Updated 6 months ago
- offical code for Dense-TSNet☆12Sep 17, 2024Updated last year
- Power-Guided Grouped SRU for Real-Time Causal Audio-Visual Speech Separation☆31Jul 20, 2026Updated last month
- ☆12Sep 12, 2024Updated last year
- ☆53Sep 10, 2024Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Source code and demo for INTERSPEECH 2024 paper: Noise-robust Speech Separation with Fast Generative Correction☆51Nov 19, 2024Updated last year
- ☆69Aug 16, 2023Updated 3 years ago
- Microphone Array Real-time System☆13Jun 7, 2017Updated 9 years ago
- Aty-TTS: Improving fairness for spoken language understanding in atypical speech with Text-to-Speech☆12May 14, 2025Updated last year
- ☆13Dec 7, 2022Updated 3 years ago
- The source code of Tim-TSENet☆15Apr 22, 2022Updated 4 years ago
- Official Implementation of TSELM: Target speaker extraction using discrete tokens and language models☆62Apr 14, 2025Updated last year
- Train no-reference speech quality estimators with multiple datasets via learned, per-dataset alignments.☆19Aug 1, 2025Updated last year
- A toolkit for researchers in the multimodal sound separation.☆16Oct 20, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆32Jan 9, 2024Updated 2 years ago
- unofficial implementation of "CPTNN: CROSS-PARALLEL TRANSFORMER NEURAL NETWORK FOR TIME-DOMAIN SPEECH ENHANCEMENT"☆15Nov 14, 2023Updated 2 years ago
- Code for paper "Noise-aware Speech Enhancement using Diffusion Probabilistic Model"☆89Jun 10, 2024Updated 2 years ago
- ☆226Dec 5, 2024Updated last year
- ☆37Jan 6, 2026Updated 7 months ago
- Implementation of the paper "Variable Bitrate Residual Vector Quantization for Audio Coding"☆11Apr 10, 2025Updated last year
- Whisper Speech Quality Assessment (WhiSQA)☆16Apr 14, 2026Updated 4 months ago
- ☆10Mar 22, 2023Updated 3 years ago
- Streaming Vocos☆33Jun 10, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official source code of the INTERSPEECH 2023 paper: "Audio-Visual Speech Separation in Noisy Environments with a Lightweight Iterative Mo…☆20Aug 20, 2026Updated 2 weeks ago
- Automatic speech annotator processing speech with voice activaty detection, overlapping speech detection, speaker diarization and automat…☆33Jun 14, 2024Updated 2 years ago
- Evaluation tool used in the BigVSAN paper☆14Mar 22, 2024Updated 2 years ago
- Unofficial implementation of wavenext vocoder☆59Aug 28, 2024Updated 2 years ago
- a Neural Vocoder supporting Ring Attention, Conformer and NSF.☆25Aug 1, 2025Updated last year
- ☆23Jul 18, 2026Updated last month
- Official code for MUSE: Flexible Voiceprint Receptive Fields and Multi-Path Fusion Enhanced Taylor Transformer for U-Net-based Speech Enh…☆58Mar 5, 2025Updated last year
- Transformer with Local Modeling by Convolution for Speech Separation and Enhancement☆136Aug 8, 2025Updated last year
- ☆45Sep 19, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Papez: Resource-Efficient Speech Separation with Auditory Working Memory (ICASSP 2023)☆22Jun 25, 2023Updated 3 years ago
- This is the official implementation of our multi-channel multi-speaker multi-spatial neural audio codec architecture.☆55Mar 17, 2025Updated last year
- A neural speech codec based on discrete WavLM representations☆26Aug 28, 2024Updated 2 years ago
- Pytorch implementation of the paper : A Global-local Attention Framework for Weakly Labelled Audio Tagging.☆13Feb 6, 2021Updated 5 years ago
- Code for the "NoiseBandNet: Controllable Time-Varying Neural Synthesis of Sound Effects Using Filterbanks" paper.☆39Jul 8, 2024Updated 2 years ago
- ☆59Apr 24, 2024Updated 2 years ago
- Code for paper Learning Audio-Visual Dereverberation☆32Aug 10, 2022Updated 4 years ago