Towards Building Text-To-Speech Systems for the Next Billion Users - Microsoft Research Intern Work - Accepted at ICASSP 2023
β57May 7, 2023Updated 3 years ago
Alternatives and similar repositories for text2speech
Users that are interested in text2speech are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Python Hindi Concatenative Based TTS using Phoneme Databaseβ26Feb 2, 2022Updated 4 years ago
- The Code shows How to Transcribe Audio to text using the fairseq_meta_mms (Google Colab Version)πβ19May 25, 2023Updated 3 years ago
- Indic-Conformer models for ASRβ22Jul 19, 2024Updated 2 years ago
- Whisper finetuned on VinBigdata-VLSP2020-100h + KenLMβ38Oct 6, 2023Updated 2 years ago
- This repository contains the HiNER dataset released with our paper at LREC 2022β17Jun 6, 2023Updated 3 years ago
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- This repository shows how to implement a basic model for multimodal entailment.β10Aug 17, 2021Updated 5 years ago
- β24Sep 1, 2023Updated 3 years ago
- On the Importance of Image Encoding in Automated Chest X-Ray Report Generation, BMVC 2022β16Dec 22, 2022Updated 3 years ago
- Text to Speech for Indic languagesβ54Mar 23, 2022Updated 4 years ago
- A Catalog lists instruction sets, models available for Indic languageβ10Mar 14, 2024Updated 2 years ago
- Prososdy Morph: A python library for manipulating pitch and duration in an algorithmic way, for resynthesizing speech.β85Jan 18, 2026Updated 7 months ago
- β45Dec 15, 2022Updated 3 years ago
- My hybrid TTS network that combines, VALL-E, VoiceBox, SpeechFlow, Seamless and TortoiseTTS into oneβ26Aug 5, 2024Updated 2 years ago
- Search any gif from Giphy APIβ12Oct 28, 2024Updated last year
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- β18Mar 16, 2023Updated 3 years ago
- Everybody Compose: Deep Beats To Musicβ12Apr 12, 2023Updated 3 years ago
- Implementation of reproducibility paper "Knowledge Graph Entity Alignment with Graph Convolutional Networks: Lessons Learned"β16Sep 28, 2020Updated 5 years ago
- Codes and datasets for our ICASSP2023 paper, Evaluating parameter-efficient transfer learning approaches on SURE benchmark for speech undβ¦β43Mar 12, 2023Updated 3 years ago
- β20Jun 4, 2026Updated 3 months ago
- The Codec 2 speech codec, compiled to WASM using Emscripten.β14Apr 27, 2023Updated 3 years ago
- Build Cordova apps with true native UIβ13May 8, 2021Updated 5 years ago
- Python code for stable matching algorithm video tutorial by The Simple Engineerβ13Nov 30, 2019Updated 6 years ago
- I built a web application using Streamlit, allowing you to make natural language queries to your database using Langchain and OpenAIβ18Jul 31, 2023Updated 3 years ago
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Python script for my article and Youtube video on building a streamlit app to use whisper for speech-to-text transcriptionβ14Mar 17, 2023Updated 3 years ago
- Text to speech is an emerging zone of AI. This repository helps to create a dataset with audio and transcripts for personalized text to sβ¦β28Mar 14, 2023Updated 3 years ago
- ezMPEG is an easy-to-use and easy-to-understand MPEG1 video encoder APIβ10Mar 26, 2017Updated 9 years ago
- Generate a UUID on all Django requests for traceabilityβ14Jul 31, 2018Updated 8 years ago
- A community driven open source alternative to Wattpadβ14Jan 9, 2023Updated 3 years ago
- Model Fusion Based Prosody Predictionβ17Mar 18, 2018Updated 8 years ago
- Code repository for the BMVC 2022 paper: Geometry Driven Progressive Warping for One Shot Face Animationβ12Jan 6, 2023Updated 3 years ago
- "Unsupervised Paraphrase Generation using Pre-trained Language Model."β22Aug 28, 2020Updated 6 years ago
- Authors official PyTorch implementation of the "HyperReenact: One-Shot Reenactment via Jointly Learning to Refine and Retarget Faces" [ICβ¦β83Sep 28, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- In app update support for cordovaβ11Oct 28, 2025Updated 10 months ago
- Source code of paper <End-to-End Language Diarization for Bilingual Code-switching Speech>β19Jan 23, 2022Updated 4 years ago
- AlphaFold2 and RoseTTAFold predictions of the SARS-CoV-2 B.1.1.529 variant Spike protein with HADDOCK antibody interactionsβ12Feb 10, 2023Updated 3 years ago
- DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism (SVS & TTS); AAAI 2022; Official codeβ10Mar 8, 2022Updated 4 years ago
- H264 encoder + MP4 output for the webβ15Dec 4, 2020Updated 5 years ago
- β13Oct 18, 2024Updated last year
- A Game Engine for J2ME Platformβ10Mar 13, 2015Updated 11 years ago