Real-time audio to chords, lyrics, beat, and melody.
☆721Aug 15, 2024Updated last year
Alternatives and similar repositories for Hybrid-Net
Users that are interested in Hybrid-Net are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Repo of the paper "Towards Building an End-to-End Multilingual Automatic Lyrics Transcription Model""☆15Jun 28, 2024Updated 2 years ago
- logWMSE, an audio quality metric with support for digital silence target. Useful for evaluating audio source separation systems, even whe…☆39Jun 24, 2025Updated last year
- Latent Space Sound Design Tool based on the VAE of stable-audio-open☆15Aug 23, 2024Updated last year
- StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models☆6,314Aug 10, 2024Updated last year
- ☆19Sep 22, 2025Updated 9 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A library for audio and music analysis, feature extraction.☆3,341Mar 6, 2026Updated 4 months ago
- All-In-One Music Structure Analyzer☆807May 9, 2024Updated 2 years ago
- text to speech☆10Mar 19, 2024Updated 2 years ago
- ☆12Nov 7, 2024Updated last year
- Self-supervised learning for real-time pitch estimation☆297Oct 15, 2025Updated 9 months ago
- A transformer-based network model for pitch detection☆166Jul 29, 2025Updated 11 months ago
- Official implementation of "Separate Anything You Describe"☆1,917Nov 26, 2024Updated last year
- Generative models for conditional audio generation☆3,820Updated this week
- Algebraic enhancements for GEMM & AI accelerators☆295Feb 28, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Music Demixing Challenge Submission Repo☆16Sep 8, 2023Updated 2 years ago
- Render waveforms to <canvas /> using WebGPU☆116Jul 3, 2026Updated 2 weeks ago
- Omniscient Mozart, being able to transcribe everything in the music, including vocal, drum, chord, beat, instruments, and more.☆1,944May 31, 2026Updated last month
- Unofficial PyTorch implementation of Music Source Separation with Band-split RNN☆191Jun 10, 2024Updated 2 years ago
- In this project we combine techniques from neural voice cloning and musical instrument synthesis to achieve good results from as little a…☆170Apr 20, 2022Updated 4 years ago
- This is the PyTorch implementation of the Universal Source Separation with Weakly labelled Data.☆368Sep 1, 2023Updated 2 years ago
- A Neural Audio Codec (NAC) for Universal Audio☆46May 30, 2025Updated last year
- WhisperFusion builds upon the capabilities of WhisperLive and WhisperSpeech to provide a seamless conversations with an AI.☆1,646Jul 31, 2024Updated last year
- a new family of super small music generation models focusing on experimental music and latent space exploration capabilities☆36May 9, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Versatile audio super resolution (any -> 48kHz) with AudioSR.☆1,926Aug 27, 2025Updated 10 months ago
- Code for ChordSync, a conformer-based audio-to-chord synchroniser☆14Oct 17, 2025Updated 9 months ago
- Algorithms as instruments.☆122Mar 2, 2025Updated last year
- Official source codes of airsep☆39Mar 26, 2024Updated 2 years ago
- C++17 port of Demucs v3 (hybrid) and v4 (hybrid transformer) models with ggml and Eigen3☆171Dec 1, 2024Updated last year
- Code for the paper Hybrid Spectrogram and Waveform Source Separation☆10,336Apr 24, 2024Updated 2 years ago
- The official implementation of DMEL the method presented in the paper "DMEL: The differentiable log-Mel spectrogram as a trainable layer …☆24Dec 21, 2024Updated last year
- End-to-end beat and downbeat tracking in the time domain.☆125Nov 10, 2021Updated 4 years ago
- Official repo of ISMIR-21 publication, “A Benchmarking Initiative for Audio-domain Music Generation using the FreeSound Loop Dataset”.☆83Nov 17, 2021Updated 4 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- An Open Source text-to-speech system built by inverting Whisper.☆4,625Dec 14, 2025Updated 7 months ago
- ☆268Feb 14, 2024Updated 2 years ago
- Pytorch implementation of the CREPE pitch tracker☆523May 16, 2025Updated last year
- Moises Source Separation Public Dataset☆191Feb 5, 2025Updated last year
- A CNN model for key estimation in music recordings☆19Aug 2, 2023Updated 2 years ago
- Daily tracking of awesome audio papers, including music generation, zero-shot tts, asr, audio generation☆409Nov 2, 2025Updated 8 months ago
- State-of-the-art audio codec with 90x compression factor. Supports 44.1kHz, 24kHz, and 16kHz mono/stereo audio.☆1,839Updated this week