The demo page of UniAudio
☆35Feb 5, 2024Updated 2 years ago
Alternatives and similar repositories for UniAudio_demo
Users that are interested in UniAudio_demo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The Open Source Code of UniAudio☆608Jul 22, 2024Updated 2 years ago
- Mustango: Toward Controllable Text-to-Music Generation☆396Jun 2, 2025Updated last year
- VoiceLDM: Text-to-Speech with Environmental Context☆194Aug 9, 2024Updated 2 years ago
- SonicVerse: Multi-Task Learning for Music Feature-Informed Captioning☆54Jul 28, 2025Updated last year
- Audiocraft is a library for audio processing and generation with deep learning. It features the state-of-the-art EnCodec audio compressor…☆20Feb 27, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- BUD-E (Buddy) is an open-source voice assistant framework that facilitates seamless interaction with AI models and APIs, enabling the cre…☆23Oct 10, 2024Updated last year
- Official repository for "Structure-Enhanced Pop Music Generation via Harmony-Aware Learning", ACM MM 2022.☆14Mar 22, 2023Updated 3 years ago
- UNMAINTAINED PROJECT☆14May 26, 2014Updated 12 years ago
- eBPF version of https://github.com/brendangregg/wss☆11Jan 26, 2023Updated 3 years ago
- Create training data for training a voice cloner for bark text to speech.☆47Jun 13, 2023Updated 3 years ago
- Score- and Lyrics-Free Singing Voice Generation☆28May 25, 2020Updated 6 years ago
- ☆25Dec 13, 2024Updated last year
- NU-Wave 2: A General Neural Audio Upsampling Model for Various Sampling Rates [WIP]☆25Jul 5, 2022Updated 4 years ago
- Colab notebook to finetune GLIDE.☆12Mar 22, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆15Sep 8, 2021Updated 5 years ago
- Simple script to re-rank images using OpenAI's CLIP https://github.com/openai/CLIP.☆15May 3, 2021Updated 5 years ago
- AutoPrep: An Automatic Preprocessing Framework for In-the-Wild Speech Data☆36Dec 31, 2023Updated 2 years ago
- Train the next generation of TTS systems.☆169Sep 13, 2024Updated 2 years ago
- The latent diffusion model for text-to-music generation.☆188Jan 26, 2024Updated 2 years ago
- Contrastive Language-Audio Pretraining☆15May 18, 2021Updated 5 years ago
- A collection of utilities for handling IPA phones.☆27Sep 24, 2023Updated 2 years ago
- MusicYOLO framework uses the object detection model, YOLOx, to locate notes in the spectrogram.☆18Jan 29, 2022Updated 4 years ago
- AudioLDM training, finetuning, evaluation and inference.☆305Dec 13, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- recent audio generation papers (including speech, music and general audios)☆13Mar 14, 2023Updated 3 years ago
- Implementation of the algorithm detailed in paper "Evolutionary design of molecules based on deep learning and a genetic algorithm"☆24Dec 15, 2023Updated 2 years ago
- ☆65Nov 4, 2021Updated 4 years ago
- Official codes and models of the paper "Auffusion: Leveraging the Power of Diffusion and Large Language Models for Text-to-Audio Generati…☆195Mar 25, 2024Updated 2 years ago
- Floral Diffusion is a custom diffusion model trained by jags using a DD 5.6 version☆26Jul 27, 2022Updated 4 years ago
- ☆27Jun 29, 2026Updated 2 months ago
- WavJourney: Compositional Audio Creation with LLMs☆544Sep 28, 2023Updated 2 years ago
- PyTorch implementation of paper "Flat Metric Minimization with Applications in Generative Modeling"☆19May 14, 2019Updated 7 years ago
- [ICML2023] Long-Term Rhythmic Video Soundtracker☆63Jul 28, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A very simple tool for situations where optimization with onnx-simplifier would exceed the Protocol Buffers upper file size limit of 2GB,…☆17Feb 24, 2026Updated 6 months ago
- ☆82Sep 8, 2026Updated last week
- Project for MIDI to Audio Synthesis☆29Mar 13, 2023Updated 3 years ago
- 🎼 text-to-video system for music visualization☆57Feb 15, 2024Updated 2 years ago
- Flow control nodes for comfyUI, allowing for more diverse workflows☆13Apr 3, 2025Updated last year
- Python code to reproduce the experiments presented in the paper Multilingual Music Genre Embeddings for Effective Cross-Lingual Music Ite…☆12Nov 13, 2020Updated 5 years ago
- simple trainer for musicgen/audiocraft☆31Jul 12, 2024Updated 2 years ago