An open source accent conversion model based on the real time voice cloning repository
☆12May 10, 2024Updated 2 years ago
Alternatives and similar repositories for accent_conversion_deep_learning
Users that are interested in accent_conversion_deep_learning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Zero-Shot Foreign Accent Conversion without a Native Reference☆36May 1, 2024Updated 2 years ago
- ☆13Jun 14, 2024Updated 2 years ago
- Resources for ACM Hack's Learn.py☆18Jun 27, 2019Updated 7 years ago
- Demo for DART, Audio Imagination workshop submission in NeurIPS 2024☆16Apr 22, 2026Updated 3 months ago
- ☆18Oct 8, 2023Updated 2 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- An unofficial PyTorch implementation of the StreamVC(Real-Time Low-Latency Voice Conversion)☆129Jun 11, 2026Updated 2 months ago
- ☆15Mar 25, 2024Updated 2 years ago
- Code for paper "Using Phonetic Posteriorgram Based Frame Pairing for Segmental Accent Conversion"☆36Jan 15, 2020Updated 6 years ago
- Run Retrieval-based Voice Conversion training and inference with ease.☆12Jan 24, 2025Updated last year
- Code for ACL 2024 main conference paper "Can We Achieve High-quality Direct Speech-to-Speech Translation Without Parallel Speech Data?".☆27Jul 2, 2024Updated 2 years ago
- ☆26Jun 5, 2024Updated 2 years ago
- ☆28Jun 22, 2026Updated last month
- A curated list of awesome AI countermeasures - tools for detecting AI-generated content☆13Jan 22, 2024Updated 2 years ago
- End-to-end Sentiment Analysis Pipeline for Call Center Conversations☆14Oct 6, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Foreign Accent Conversion by Synthesizing Speech from Phonetic Posteriorgrams (Interspeech'19)☆147Jul 6, 2023Updated 3 years ago
- A sequence-to-sequence voice conversion toolkit.☆113Mar 15, 2026Updated 4 months ago
- This repository includes the code to reproduce our paper Partially-Connected Differentiable Architecture Search for Deepfake and Spoofing…☆18Apr 30, 2022Updated 4 years ago
- text to speech☆10Mar 19, 2024Updated 2 years ago
- Differentiable Mean Opinion Score Regularization for Perceptual Speech Enhancement☆25Apr 16, 2023Updated 3 years ago
- A TensorFlow implementation of a variational autoencoder-generative adversarial network (VAE-GAN) architecture for speech-to-speech style…☆19Jan 17, 2022Updated 4 years ago
- A real-time voice conversion model based on VITS.☆16Aug 1, 2024Updated 2 years ago
- Koa.js framework setup to run within Next.js API routes.☆11Jul 27, 2026Updated 2 weeks ago
- Implementation of the paper "Variable Bitrate Residual Vector Quantization for Audio Coding"☆11Apr 10, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- A Python wrapper for libhackrf☆12Jul 10, 2023Updated 3 years ago
- Parallel Prefix Sum (Scan) with CUDA☆30Jun 22, 2024Updated 2 years ago
- 🫠 check your data, before you wreck your model☆16Aug 11, 2022Updated 3 years ago
- Creates video from TTS output and viseme images.☆16Jun 18, 2022Updated 4 years ago
- ☆11Mar 7, 2025Updated last year
- ☆13Nov 14, 2024Updated last year
- a mindful website blocker for the productive.☆66Apr 10, 2025Updated last year
- ☆11Jul 21, 2024Updated 2 years ago
- Chinese entity relation extraction☆21Apr 26, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- PromptTTS++: Controlling Speaker Identity in Prompt-Based Text-To-Speech Using Natural Language Descriptions☆86Oct 11, 2024Updated last year
- PyTorch Implementation of ViT-TTS (EMNLP'23)☆11Oct 20, 2023Updated 2 years ago
- ☆11Feb 20, 2025Updated last year
- Swipe Right On A New Peering Relationship☆15Jun 21, 2020Updated 6 years ago
- A deepfake audio dataset for detecting fake speech from codec-based speech synthesis systems, Interspeech 2024☆22Jul 27, 2024Updated 2 years ago
- mobile Driving License Interoperability Learning Platform☆15Sep 6, 2018Updated 7 years ago
- [CVPR'24 Oral] Metacloak: Preventing Unauthorized Subject-driven Text-to-image Diffusion-based Synthesis via Meta-learning☆32Nov 19, 2024Updated last year