Conditioned U-Net for Music Source Separation
☆20May 15, 2021Updated 5 years ago
Alternatives and similar repositories for conditioned-u-net
Users that are interested in conditioned-u-net are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation of A cappella: Audio-visual Singing VoiceSeparation, from BMVC21☆18May 14, 2022Updated 4 years ago
- A PyTorch Implementation of the paper - Choi, Woosung, et al. "Investigating u-nets with various intermediate blocks for spectrogram-base…☆80Jul 1, 2022Updated 4 years ago
- Implementations for master thesis "Musical Instrument Recognition in Multi-Instrument Audio Contexts" with MedleyDB.☆16Apr 4, 2019Updated 7 years ago
- ☆17Aug 18, 2022Updated 3 years ago
- ☆14Nov 22, 2022Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Voice Framework☆18Jan 21, 2026Updated 6 months ago
- ☆26Mar 5, 2018Updated 8 years ago
- Block-Online Multi-Channel Speech Enhancement Using DNN-Supported Relative Transfer Function Estimates☆34May 26, 2020Updated 6 years ago
- ☆11Oct 14, 2020Updated 5 years ago
- Room impulse response simulation for various array architectures using Monte-Carlo simulation and quaternions (Python)☆18Feb 25, 2026Updated 5 months ago
- Visually-informed Music Source Separation project at Jeju 2018 Deep Learning Summer Camp☆30Sep 14, 2018Updated 7 years ago
- A Benchmark and Evaluation Suite for Zero-shot Singing Voice Synthesis☆33Feb 11, 2026Updated 5 months ago
- FastLongSpeech is a novel framework designed to extend the capabilities of Large Speech-Language Models for efficient long-speech process…☆16Jul 22, 2025Updated last year
- A PyTorch implementation of Meta-TasNet from "Meta-learning Extractors for Music Source Separation☆138Jul 25, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆12Oct 9, 2025Updated 9 months ago
- ☆19Aug 23, 2024Updated last year
- ☆13Nov 26, 2019Updated 6 years ago
- A simple and humble image captioning application, based on a neural network built with Keras☆10Sep 23, 2022Updated 3 years ago
- Glow-TTS with Stochastic Duration Predictor and Stochastic Pitch Predictor☆19Jun 5, 2023Updated 3 years ago
- Supplementary code for the experiments described in the 2021 ISMIR submission: Leveraging Hierarchical Structures for Few Shot Musical In…☆41Aug 12, 2022Updated 3 years ago
- Code for the ISMIR 2021 tutorial "Programming MIR Baselines from Scratch: Three Cases Studies"☆30Nov 21, 2021Updated 4 years ago
- Pytorch implementation of Deepmind's WaveRNN model☆13Apr 5, 2020Updated 6 years ago
- Event Relation in Text-to-Audio (TTA) Generation☆21Feb 26, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Speech Resynthesis and Language Modeling☆27Jun 11, 2025Updated last year
- A Tensorflow LSTM spam detector utilizing GloVe word embeddings.☆12Nov 9, 2019Updated 6 years ago
- ISMIR2018 Tutorial on Open Source and Reproducibility in MIR Research☆41Aug 16, 2022Updated 3 years ago
- Self-supervised VQ-VAE for One-Shot Music Style Transfer☆99Feb 24, 2025Updated last year
- [NeurIPS 2023 - ML for Audio Workshop (Oral)] Zero-shot audio captioning with audio-language model guidance and audio context keywords☆19Nov 30, 2024Updated last year
- Code for "A diffusion-inspired training strategy for singing voice extraction in the waveform domain" (ISMIR 2022)☆17Feb 16, 2023Updated 3 years ago
- This is the official repository of ``Scalable Neural Vocoder from Range-Null Space Decomposition'', which is submitted to TPAMI.☆54Oct 11, 2025Updated 9 months ago
- ☆15Sep 26, 2022Updated 3 years ago
- Training and evaluation code for Re-MOVE models with embedding distillation☆31Jul 6, 2023Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- HpTF compensation filters for binaural synthesis with Matlab script for creation of filters out of measurement data☆16Mar 21, 2017Updated 9 years ago
- ☆16May 28, 2018Updated 8 years ago
- FCTalker: Fine and Coarse Grained Context Modeling for Expressive Conversational Speech Synthesis (Accepted by ISCSLP'2024)☆26Feb 22, 2024Updated 2 years ago
- Official Repository for "SingFake: Singing Voice Deepfake Detection"☆64Feb 26, 2024Updated 2 years ago
- pytorch implementation for MultiSpeech: Multi-Speaker Text to Speech with Transformer paper☆21Jun 23, 2022Updated 4 years ago
- Subband system identification using generalized Weighted Overlap-Add (WOLA) filter bank for improved acoustic echo cancellation.☆15May 8, 2025Updated last year
- ☆27May 14, 2020Updated 6 years ago