Spatial Voice Conversion: Voice Conversion Preserving Spatial Information and Non-target Signals
☆18Aug 8, 2024Updated 2 years ago
Alternatives and similar repositories for spatial_voice_conversion
Users that are interested in spatial_voice_conversion are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A fully and partially fake speech dataset for evaluation☆15Nov 11, 2025Updated 9 months ago
- ☆17Dec 18, 2023Updated 2 years ago
- ☆27Aug 2, 2024Updated 2 years ago
- SS2 HRTF Dataset - Reality Labs Research Audio☆18May 22, 2026Updated 3 months ago
- A command-line tool that provides the core functionality for storing and retrieving shell command history with directory context in SQLit…☆11Aug 8, 2026Updated 3 weeks ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Official repository for Mamba-based Segmentation Model for Speaker Diarization☆47May 13, 2025Updated last year
- ☆11May 7, 2022Updated 4 years ago
- Code for ACL 2024 main conference paper "Can We Achieve High-quality Direct Speech-to-Speech Translation Without Parallel Speech Data?".☆27Jul 2, 2024Updated 2 years ago
- Collect eye movement data using a webcam(with calibration).☆10Mar 25, 2021Updated 5 years ago
- Named entity recognition for scientific and vernacular plant names☆14Jan 17, 2023Updated 3 years ago
- ChatGPT for Scratch☆18Mar 8, 2026Updated 5 months ago
- PyTorch implementation of Swin Transformer for 1-dimensional data☆19Mar 15, 2024Updated 2 years ago
- Official implementation of "Wave-Trainer-Fit: Neural Vocoder with Trainable Prior and Fixed-Point Iteration towards High-Quality Speech G…☆16Feb 6, 2026Updated 6 months ago
- Glow-TTS with Stochastic Duration Predictor and Stochastic Pitch Predictor☆19Jun 5, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This is a song listening and music recognition project based on audio fingerprint algorithm.☆11Mar 26, 2022Updated 4 years ago
- ☆28Jun 22, 2026Updated 2 months ago
- ☆40Jan 24, 2023Updated 3 years ago
- Just another FastSpeech 2 but cleaner code :)☆29Jun 28, 2024Updated 2 years ago
- Code for the paper "Toward Fully Self-Supervised Multi-Pitch Estimation".☆25Sep 27, 2025Updated 11 months ago
- Unofficial Pytorch implementation of SNAC: Speaker-normalized affine coupling layer in flow-based architecture for zero-shot multi-speake…☆57Aug 7, 2023Updated 3 years ago
- A lightweight audio codec based on a single quantizer☆72Aug 15, 2025Updated last year
- ☆45Sep 19, 2024Updated last year
- ☆11Apr 20, 2020Updated 6 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆69Aug 16, 2023Updated 3 years ago
- Digital Speech Processing in PyTorch.☆15Aug 12, 2022Updated 4 years ago
- ☆28Apr 24, 2026Updated 4 months ago
- Demucs Lightning: A PyTorch lightning version of Demucs with Hydra and Tensorboard features☆85May 3, 2023Updated 3 years ago
- Template Code for TOC Project 2017☆10Apr 26, 2017Updated 9 years ago
- Generator for anechoic, non-stationary noise signals☆12Aug 12, 2022Updated 4 years ago
- ☆14Aug 19, 2024Updated 2 years ago
- ☆41May 15, 2023Updated 3 years ago
- ☆12Apr 1, 2020Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆14Feb 3, 2026Updated 6 months ago
- Official Repository for "GOTCHA: Real-Time Video Deepfake Detection via Challenge-Response"☆11Jul 8, 2024Updated 2 years ago
- Directional sparse filtering for blind speech separation☆11Jun 8, 2021Updated 5 years ago
- A query by humming system based on locality sensitive hashing indexes☆12May 8, 2014Updated 12 years ago
- 🎙️ Automatically transcribe audio/video into high-quality, speaker-specific Text-To-Speech datasets ✨☆18May 20, 2025Updated last year
- SpeechFake: A Large-Scale Multilingual Speech Deepfake Dataset Incorporating Cutting-Edge Generation Methods☆27Aug 13, 2025Updated last year
- Interface Design for Self-Supervised Speech Models, Accepted to Interspeech2024☆16Nov 19, 2024Updated last year