LLaSE: Maximizing Acoustic Preservation for LLaMA based Speech Enhancement
☆16Jul 11, 2025Updated last year
Alternatives and similar repositories for LLaSE
Users that are interested in LLaSE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- LLaSE-G1: Incentivizing Generalization Capability for LLaMA-based Speech Enhancement☆105Apr 1, 2025Updated last year
- ☆17Apr 9, 2026Updated 3 months ago
- Reimplementation of Miipher☆30Aug 16, 2023Updated 2 years ago
- LLaSE-G1: Incentivizing Generalization Capability for LLaMA-based Speech Enhancement☆47Mar 10, 2025Updated last year
- FINALLY: Fast and universal speech enhancement model delivering studio-quality audio for a wide range of recordings.☆28Apr 1, 2026Updated 3 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- We propose C2SER, a novel audio-language model designed to enhance the stability and accuracy of speech emotion recognition through conte…☆49Mar 3, 2025Updated last year
- Generation scripts for EARS-WHAM and EARS-Reverb☆48Jul 4, 2025Updated last year
- Voice Framework☆18Jan 21, 2026Updated 6 months ago
- Personalized AEC☆19Nov 3, 2022Updated 3 years ago
- A solution to denoising and separating for two-speaker-mixed noisy speech, using a BSRNN inspired network.☆15Aug 22, 2023Updated 2 years ago
- Dataset simulation for DPCCN.☆16Dec 25, 2022Updated 3 years ago
- ☆33Jun 1, 2026Updated last month
- ☆36Apr 14, 2026Updated 3 months ago
- Production first, nn-based on-device signal processing toolkit.☆63May 30, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Fully quantized Neural Networks for Audio Source Separation☆17Aug 11, 2024Updated last year
- Python scripts to create noisy and reverberant 2-speaker mixture audio with Libri-Light and WHAM☆17Nov 7, 2024Updated last year
- ☆17Jun 2, 2025Updated last year
- ☆25Feb 28, 2023Updated 3 years ago
- Room impulse response simulation for various array architectures using Monte-Carlo simulation and quaternions (Python)☆18Feb 25, 2026Updated 4 months ago
- Apply Score diffusion to improve speech signals recorded under various adverse conditions and distortions, including noise, reverberation…☆82Jul 29, 2024Updated last year
- My hybrid TTS network that combines, VALL-E, VoiceBox, SpeechFlow, Seamless and TortoiseTTS into one☆26Aug 5, 2024Updated last year
- The source code of Tim-TSENet☆15Apr 22, 2022Updated 4 years ago
- offical code for Dense-TSNet☆12Sep 17, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆43Feb 8, 2025Updated last year
- PyTorch implementation of TinyWASE described in our paper "Compressing Speaker Extraction Model with Ultra-low Precision Quantization and…☆11Jun 28, 2021Updated 5 years ago
- Audio-FLAN☆161Sep 23, 2025Updated 10 months ago
- [Early Alpha] A unified framework for text-to-speech, voice conversion, automatic speech recognition, audio classification, voice activit…☆22Jan 10, 2025Updated last year
- Joint magnitude estimation and phase recovery using Cycle-in-Cycle GAN for non-parallel speech enhancement☆10Jan 24, 2022Updated 4 years ago
- ☆25Jan 24, 2023Updated 3 years ago
- ☆39Feb 23, 2022Updated 4 years ago
- PyTorch-based room impulse response (RIR) simulation toolkit with dynamic scenes, GPU acceleration.☆22Updated this week
- Subband system identification using generalized Weighted Overlap-Add (WOLA) filter bank for improved acoustic echo cancellation.☆15May 8, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Data simulation scripts for paper "Target Sound Extraction with Variable Cross-modality Clues"☆17May 19, 2023Updated 3 years ago
- Ablation study of local spectral attention (LSA) for full-band speech enhancement (SE)☆28Sep 16, 2023Updated 2 years ago
- ☆33Nov 29, 2022Updated 3 years ago
- An unofficial implementation of the Personal VAD speaker-conditioned voice activity detection method. Bachelor's thesis project.☆90Sep 22, 2022Updated 3 years ago
- ☆26Mar 31, 2026Updated 3 months ago
- ☆62Apr 11, 2022Updated 4 years ago
- ☆11Jun 6, 2022Updated 4 years ago