☆19Jun 29, 2026Updated last month
Alternatives and similar repositories for LeVo
Users that are interested in LeVo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ESLTTS dataset☆16Feb 6, 2025Updated last year
- Evaluation tool used in the BigVSAN paper☆14Mar 22, 2024Updated 2 years ago
- ☆37Mar 26, 2024Updated 2 years ago
- TriNet: stabilizing self-supervised learning from complete or slow collapse on ASR.☆34Jun 1, 2023Updated 3 years ago
- ☆21Jul 16, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆20Jun 5, 2022Updated 4 years ago
- We design a spectral compression mapping (SCM) for full-band speech enhancement, and propose a two-stage stream named MHA-DPCRN☆24Jul 4, 2022Updated 4 years ago
- ☆20Jul 13, 2022Updated 4 years ago
- ☆10Jun 11, 2024Updated 2 years ago
- Official Implementation of "Colored Noise Diffusion Sampling"☆39Jun 1, 2026Updated 2 months ago
- Pytorch implementation for “V2C: Visual Voice Cloning”☆35Jan 28, 2023Updated 3 years ago
- A pitch detection model trained to be robust against noise and reverberation environments.☆27Jan 21, 2025Updated last year
- Updated folk of g2pk☆13Aug 18, 2023Updated 2 years ago
- Implementation of SpatialCodec.☆71Sep 23, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆36Aug 21, 2021Updated 4 years ago
- PyTorch Implementation of TCSinger 2(ACL 2025): Customizable Multilingual Zero-shot Singing Voice Synthesis☆182Apr 19, 2026Updated 3 months ago
- Hybrid Flow Matching and GAN with Multi-Resolution Network for Few-Step High-Fidelity Audio Generation☆146Mar 8, 2026Updated 4 months ago
- MusicYOLO framework uses the object detection model, YOLOx, to locate notes in the spectrogram.☆18Jan 29, 2022Updated 4 years ago
- ☆35Oct 23, 2025Updated 9 months ago
- Kling-Foley: Multimodal Diffusion Transformer for High-Quality Video-to-Audio Generation☆62Jun 26, 2025Updated last year
- A jigsaw puzzle solver using AI☆14May 18, 2026Updated 2 months ago
- GPT-style network for phonemization with durations of text☆68Mar 21, 2024Updated 2 years ago
- 2020년 21대 국회의원 총선거 지도☆11Mar 19, 2020Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆13Mar 23, 2026Updated 4 months ago
- Synthesized singing voice demos of WeSinger 2 paper.☆26Feb 20, 2023Updated 3 years ago
- Music -- separate into stems, modify, convert to midi, add synth sounds, remix.☆20Updated this week
- Implementation of Google's USM speech model in Pytorch☆36Jul 27, 2026Updated last week
- Official implementation of "Automatic Tuning of Loss Trade-offs without Hyper-parameter Search in End-to-End Zero-Shot Speech Synthesis",…☆80May 29, 2023Updated 3 years ago
- VISinger 2: High-Fidelity End-to-End Singing Voice Synthesis Enhanced by Digital Signal Processing Synthesizer☆356Nov 4, 2024Updated last year
- DACVAE☆227Dec 22, 2025Updated 7 months ago
- Official implementation of "Avocodo: Generative Adversarial Network for Artifact-Free Vocoder" (AAAI2023)☆154Feb 1, 2023Updated 3 years ago
- A robust pitch tracker using synchro-squeezed fft and frequency domain autocorrelation☆38Jan 17, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- StreamAvatar: Streaming Diffusion Models for Real-Time Interactive Human Avatars☆18Mar 31, 2026Updated 4 months ago
- (WIP)long form speech generatoins☆30Apr 2, 2025Updated last year
- ☆17May 1, 2026Updated 3 months ago
- Code repository for paper "Tuning Audio Diffusion Models through Activation Steering"☆21May 27, 2026Updated 2 months ago
- Official Implement of Multi-Stage Multi-Codebook (MSMC) TTS☆168Apr 10, 2024Updated 2 years ago
- ☆88Nov 1, 2022Updated 3 years ago
- 语智科技远场(单麦克风)语音识别引擎 FFASR 接入指南☆15Aug 4, 2023Updated 2 years ago