Create training data for training a voice cloner for bark text to speech.
β47Jun 13, 2023Updated 3 years ago
Alternatives and similar repositories for bark-data-gen
Users that are interested in bark-data-gen are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The code for the bark-voicecloning model. Training and inference.β710Sep 13, 2023Updated 2 years ago
- π Text-prompted Generative Audio Model - With the ability to clone voicesβ21May 17, 2023Updated 3 years ago
- Official repository for "Structure-Enhanced Pop Music Generation via Harmony-Aware Learning", ACM MM 2022.β14Mar 22, 2023Updated 3 years ago
- [Batching/MultiGPU/DataLoader Implemented] Code for the paper Hybrid Spectrogram and Waveform Source Separationβ24Aug 2, 2023Updated 3 years ago
- β18Jan 20, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Montreal Forced Aligner for Vietnameseβ15Oct 23, 2023Updated 2 years ago
- Audio generation using diffusion models, in PyTorch.β49Sep 28, 2023Updated 2 years ago
- Community-controlled voice data collection for language preservation and AI development. Companion to 'AI Techniques for Indigenous Cultuβ¦β71May 6, 2026Updated 3 months ago
- β63Jan 15, 2024Updated 2 years ago
- SLMGAN: Exploiting Speech Language Model Representations for Unsupervised Zero-Shot Voice Conversion in GANsβ16Jul 19, 2023Updated 3 years ago
- This is the code and dataset repo for Interspeech 2024 paper "Target conversation extraction: Source separation using turn-taking dynamicβ¦β59Aug 15, 2025Updated last year
- FreeVC: Towards High-Quality Text-Free One-Shot Voice Conversionβ716Jan 19, 2025Updated last year
- Site for sharing Bark voicesβ50Jul 6, 2026Updated last month
- Implementation of AudioLM, a SOTA Language Modeling Approach to Audio Generation out of Google Research, in Pytorchβ2,626Jan 12, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- β13May 23, 2024Updated 2 years ago
- Image inpainting system frontendβ14Jan 31, 2023Updated 3 years ago
- Google's TPGST reimplementation.β34Dec 11, 2019Updated 6 years ago
- Use VITS and Opencpop to develop singing voice synthesis; Different from VISinger.β38Feb 24, 2023Updated 3 years ago
- Make-A-Video Latent Diffusion Modelβ19Nov 15, 2023Updated 2 years ago
- Findings of ACL 2023 | AlignSTS: a speech-to-singing (STS) model based on modality disentanglement and cross-modal alignmentβ67Jul 5, 2024Updated 2 years ago
- Russian open TTS datasetβ18Nov 5, 2019Updated 6 years ago
- Barkify: an unoffical training implementation of Bark TTS by suno-aiβ130May 31, 2023Updated 3 years ago
- Codebase for ICLR' 23 paper- ''wav2tok: Deep Sequence Tokenizer for Audio Retrieval"β37Jun 30, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [IEEE/CVF CVPR'2022] "ST-MFNet: A Spatio-Temporal Multi-Flow Network for Frame Interpolation", Duolikun Danier, Fan Zhang, David Bullβ13Oct 9, 2023Updated 2 years ago
- π Create labeled datasets, enhance audio quality, identify speakers, support diverse dataset types. π§π₯π Advanced audio processing.β262Jun 10, 2024Updated 2 years ago
- Voice Conversion method based on speaker styleβ14Aug 7, 2021Updated 5 years ago
- Vietnamese Voice Cloning System using Speaker Verification training on multispeaker VITSβ56Dec 1, 2023Updated 2 years ago
- Web Audio API Node Editorβ16Mar 7, 2023Updated 3 years ago
- A silly and weirdly useful experiment where I attempt to encode one bit of information with a VAEβ11Dec 31, 2016Updated 9 years ago
- Remove generated stories with stray unicode charactersβ12Jan 3, 2024Updated 2 years ago
- Official source codes of airsepβ39Mar 26, 2024Updated 2 years ago
- Everybody Compose: Deep Beats To Musicβ12Apr 12, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- OpenOrca-KO datasetμ νμ©νμ¬ llama2λ₯Ό fine-tuningν Korean-OpenOrcaβ18Nov 1, 2023Updated 2 years ago
- Implementation of Natural Speech 2, Zero-shot Speech and Singing Synthesizer, in Pytorchβ1,333Sep 24, 2023Updated 2 years ago
- Unofficial implementation JEN-1 Composer: A Unified Framework for High-Fidelity Multi-Track Music Generation(https://arxiv.org/abs/2310.1β¦β32Jan 19, 2024Updated 2 years ago
- Voice data <= 10 mins can also be used to train a good VC model!β12Dec 5, 2023Updated 2 years ago
- Basic framework for training Dreambooth Stable Diffusion v1.5 on Banana's v1.0 serverless GPU platformβ37Nov 15, 2022Updated 3 years ago
- Vector-Quantized Contrastive Predictive Coding for Acoustic Unit Discovery and Voice Conversionβ142Sep 1, 2020Updated 6 years ago
- β69May 19, 2023Updated 3 years ago