The demo page of UniAudio
☆35Feb 5, 2024Updated 2 years ago
Alternatives and similar repositories for UniAudio_demo
Users that are interested in UniAudio_demo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The Open Source Code of UniAudio☆608Jul 22, 2024Updated 2 years ago
- A solution to denoising and separating for two-speaker-mixed noisy speech, using a BSRNN inspired network.☆15Aug 22, 2023Updated 3 years ago
- Mustango: Toward Controllable Text-to-Music Generation☆395Jun 2, 2025Updated last year
- VoiceLDM: Text-to-Speech with Environmental Context☆194Aug 9, 2024Updated 2 years ago
- SonicVerse: Multi-Task Learning for Music Feature-Informed Captioning☆54Jul 28, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Chorale Music Separation Dataset and Model Framework☆41Dec 5, 2022Updated 3 years ago
- Audiocraft is a library for audio processing and generation with deep learning. It features the state-of-the-art EnCodec audio compressor…☆20Feb 27, 2024Updated 2 years ago
- BUD-E (Buddy) is an open-source voice assistant framework that facilitates seamless interaction with AI models and APIs, enabling the cre…☆23Oct 10, 2024Updated last year
- Official repository for "Structure-Enhanced Pop Music Generation via Harmony-Aware Learning", ACM MM 2022.☆14Mar 22, 2023Updated 3 years ago
- UNMAINTAINED PROJECT☆14May 26, 2014Updated 12 years ago
- eBPF version of https://github.com/brendangregg/wss☆11Jan 26, 2023Updated 3 years ago
- Create training data for training a voice cloner for bark text to speech.☆47Jun 13, 2023Updated 3 years ago
- Score- and Lyrics-Free Singing Voice Generation☆28May 25, 2020Updated 6 years ago
- ☆193Aug 23, 2026Updated last week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆25Dec 13, 2024Updated last year
- NU-Wave 2: A General Neural Audio Upsampling Model for Various Sampling Rates [WIP]☆25Jul 5, 2022Updated 4 years ago
- ☆15Sep 8, 2021Updated 4 years ago
- Simple script to re-rank images using OpenAI's CLIP https://github.com/openai/CLIP.☆15May 3, 2021Updated 5 years ago
- AutoPrep: An Automatic Preprocessing Framework for In-the-Wild Speech Data☆36Dec 31, 2023Updated 2 years ago
- Train the next generation of TTS systems.☆169Sep 13, 2024Updated last year
- Unified API to facilitate usage of pre-trained "perceptor" models, a la CLIP☆39Nov 26, 2022Updated 3 years ago
- Continuous descriptor-based control for deep audio synthesis☆23Aug 4, 2023Updated 3 years ago
- A collection of utilities for handling IPA phones.☆27Sep 24, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official repository for the paper "AudioMAE++: learning better masked audio representations with SwiGLU FFNs"☆15Apr 30, 2026Updated 4 months ago
- MusicYOLO framework uses the object detection model, YOLOx, to locate notes in the spectrogram.☆18Jan 29, 2022Updated 4 years ago
- AudioLDM training, finetuning, evaluation and inference.☆305Dec 13, 2024Updated last year
- recent audio generation papers (including speech, music and general audios)☆13Mar 14, 2023Updated 3 years ago
- ☆65Nov 4, 2021Updated 4 years ago
- Floral Diffusion is a custom diffusion model trained by jags using a DD 5.6 version☆26Jul 27, 2022Updated 4 years ago
- ☆25Jun 29, 2026Updated 2 months ago
- WavJourney: Compositional Audio Creation with LLMs☆544Sep 28, 2023Updated 2 years ago
- PyTorch implementation of paper "Flat Metric Minimization with Applications in Generative Modeling"☆19May 14, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICML2023] Long-Term Rhythmic Video Soundtracker☆63Jul 28, 2025Updated last year
- ☆82Jan 22, 2025Updated last year
- Official code for the paper "Compositional Generalization from First Principles" (NeurIPS 2023)☆15Jul 25, 2023Updated 3 years ago
- Project for MIDI to Audio Synthesis☆29Mar 13, 2023Updated 3 years ago
- This setup allows to train end-to-end neural models for spoken language understanding (SLU).☆11Jun 12, 2023Updated 3 years ago
- 🎼 text-to-video system for music visualization☆57Feb 15, 2024Updated 2 years ago
- Flow control nodes for comfyUI, allowing for more diverse workflows☆13Apr 3, 2025Updated last year