Towards Building Text-To-Speech Systems for the Next Billion Users - Microsoft Research Intern Work - Accepted at ICASSP 2023
β57May 7, 2023Updated 3 years ago
Alternatives and similar repositories for text2speech
Users that are interested in text2speech are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- EC499: Major Projectβ11Jun 25, 2023Updated 3 years ago
- The Code shows How to Transcribe Audio to text using the fairseq_meta_mms (Google Colab Version)πβ19May 25, 2023Updated 3 years ago
- Indic-Conformer models for ASRβ19Jul 19, 2024Updated 2 years ago
- Text-to-Speech for languages of Indiaβ378Nov 8, 2024Updated last year
- Whisper finetuned on VinBigdata-VLSP2020-100h + KenLMβ38Oct 6, 2023Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- β11Oct 22, 2023Updated 2 years ago
- SaveRestrictedContentBot @AM_ROBOTSβ11Oct 29, 2022Updated 3 years ago
- Text to Speech for Indic languagesβ53Mar 23, 2022Updated 4 years ago
- A Catalog lists instruction sets, models available for Indic languageβ10Mar 14, 2024Updated 2 years ago
- AI and IoT based Smart Parkingβ10Apr 15, 2022Updated 4 years ago
- A JavaScript and TypeScript port of PyTorch C++ library (libtorch) - Node.js N-API bindings for libtorch.β17Jan 15, 2023Updated 3 years ago
- Transformer-based Model to recognize any of 7 unique intents from the Snips personal voice assistant.β13Mar 11, 2022Updated 4 years ago
- My hybrid TTS network that combines, VALL-E, VoiceBox, SpeechFlow, Seamless and TortoiseTTS into oneβ26Aug 5, 2024Updated last year
- Official implementation of the paper "Evading Forensic Classifiers with Attribute-Conditioned Adversarial Faces" (CVPR 23)β46Jan 24, 2024Updated 2 years ago
- Open source password manager - Proton Pass β’ AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Audio GIFs (.a.gif) -- "Sounds like a Bad Idea."β12Jun 7, 2019Updated 7 years ago
- Search any gif from Giphy APIβ12Oct 28, 2024Updated last year
- Everybody Compose: Deep Beats To Musicβ12Apr 12, 2023Updated 3 years ago
- Implementation of reproducibility paper "Knowledge Graph Entity Alignment with Graph Convolutional Networks: Lessons Learned"β16Sep 28, 2020Updated 5 years ago
- Codes and datasets for our ICASSP2023 paper, Evaluating parameter-efficient transfer learning approaches on SURE benchmark for speech undβ¦β43Mar 12, 2023Updated 3 years ago
- The Codec 2 speech codec, compiled to WASM using Emscripten.β14Apr 27, 2023Updated 3 years ago
- β20Jun 4, 2026Updated last month
- Express.js ported to a Service Worker contextβ18Mar 6, 2025Updated last year
- A simple, yet effective, walking AI.β12Aug 30, 2020Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Deep learning-based audio spoofing attack detection experiments for speaker verification.β14Apr 20, 2023Updated 3 years ago
- Audio classification using Machine Learningβ13Dec 17, 2015Updated 10 years ago
- Build Cordova apps with true native UIβ13May 8, 2021Updated 5 years ago
- An easy-to-use, efficient, powerful tool to remote update your cordova app.β11Aug 13, 2018Updated 7 years ago
- Python script for my article and Youtube video on building a streamlit app to use whisper for speech-to-text transcriptionβ15Mar 17, 2023Updated 3 years ago
- Text to speech is an emerging zone of AI. This repository helps to create a dataset with audio and transcripts for personalized text to sβ¦β28Mar 14, 2023Updated 3 years ago
- β22Jul 11, 2023Updated 3 years ago
- Detection and Classification of UI Elements of Web pages and Apps from Wireframe Sketchesβ10Oct 9, 2023Updated 2 years ago
- ezMPEG is an easy-to-use and easy-to-understand MPEG1 video encoder APIβ11Mar 26, 2017Updated 9 years ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- β32Dec 4, 2022Updated 3 years ago
- Model Fusion Based Prosody Predictionβ17Mar 18, 2018Updated 8 years ago
- Code repository for the BMVC 2022 paper: Geometry Driven Progressive Warping for One Shot Face Animationβ12Jan 6, 2023Updated 3 years ago
- Pretraining, fine-tuning and evaluation scripts for Indic-Wav2Vec2β117Aug 28, 2025Updated 11 months ago
- "Unsupervised Paraphrase Generation using Pre-trained Language Model."β22Aug 28, 2020Updated 5 years ago
- In app update support for cordovaβ11Oct 28, 2025Updated 9 months ago
- Source code of paper <End-to-End Language Diarization for Bilingual Code-switching Speech>β19Jan 23, 2022Updated 4 years ago