An HTML interface for finetuning the sync map output from aeneas
☆53Jul 5, 2022Updated 4 years ago
Alternatives and similar repositories for finetuneas
Users that are interested in finetuneas are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Postprocess SRT derived speech alignments for creating clean datasets for machine learning☆17Jan 4, 2023Updated 3 years ago
- lachesis automates the segmentation of a transcript into closed captions☆35Jan 26, 2017Updated 9 years ago
- PyTorch speech2text inference script for the NVidia openseq2seq wav2letter model variant☆10Aug 12, 2019Updated 7 years ago
- ☆23Jul 22, 2022Updated 4 years ago
- My public domain speech index☆13Sep 19, 2019Updated 7 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- aeneas is a Python/C library and a set of tools to automagically synchronize audio and text (aka forced alignment)☆2,865Jul 25, 2026Updated last month
- Open Source Crimean Tatar Text-to-Speech datasets☆14Feb 23, 2025Updated last year
- Jupyter Notebooks for creating Speech datasets☆46Mar 3, 2019Updated 7 years ago
- TTS Client for Coqui TTS server☆13Jan 7, 2023Updated 3 years ago
- End-to-end Text-to-Speech with Generative Adversarial Networks☆20Feb 6, 2021Updated 5 years ago
- Evaluation of STT models for german language☆16Jan 22, 2022Updated 4 years ago
- Evaluate results from ASR/Speech-to-Text quickly☆41Dec 28, 2021Updated 4 years ago
- Fast trigram-indexed regex search for codebases — 2-6x faster than ripgrep☆21Mar 24, 2026Updated 5 months ago
- This repo contains the baseline model recipes and pre-trained model for GramVanni hindi ASR challenge☆16Mar 26, 2022Updated 4 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Running Mozilla's implementation of Baidu DeepSpeech on Google Colaboratory☆16Mar 18, 2019Updated 7 years ago
- ☆14Mar 31, 2023Updated 3 years ago
- Web-based tool for correcting speech-to-text generated transcripts of oral histories.☆10Mar 26, 2026Updated 5 months ago
- Repo for NYPL's 2016 Event, Open Audio Weekend☆14Jun 30, 2016Updated 10 years ago
- FFTNet vocoder implementation☆81Sep 28, 2018Updated 7 years ago
- ☆10Apr 8, 2024Updated 2 years ago
- NISQA - Non-Intrusive Speech Quality and TTS Naturalness Assessment☆16Apr 13, 2022Updated 4 years ago
- Trained speaker embedding deep learning models and evaluation pipelines in pytorch and tesorflow for speaker recognition.☆36Oct 4, 2019Updated 6 years ago
- Sequence-to-sequence TTS based on Kyubyong's dc_tts☆61Feb 2, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- This Guidance demonstrates how to validate checksums for compliance and audit requirements with an on-demand fixity check process.☆14Oct 20, 2024Updated last year
- The YouTube Text-To-Speech dataset is comprised of waveform audio extracted from YouTube videos alongside their English transcriptions☆53Apr 1, 2021Updated 5 years ago
- AWS Transcribe evaluation pipeline: bulk-process audio files and view the results☆17Oct 13, 2023Updated 2 years ago
- Проект для перевода чисел, записанных в текстовом виде на русском языке.☆11Apr 5, 2022Updated 4 years ago
- Speech-to-text for podcasting prototypes☆16Jan 19, 2013Updated 13 years ago
- Create a custom Watson Speech to Text model using specialized domain data☆62Aug 31, 2021Updated 5 years ago
- PAVOQUE Corpus of Expressive Speech☆12Aug 2, 2016Updated 10 years ago
- My guide to create an italian TTS with Coqui☆14Feb 2, 2022Updated 4 years ago
- Docker Image for Low-cost HD surveillance Camera Module on Raspberry Pi 3☆21Jun 3, 2020Updated 6 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆15Oct 11, 2019Updated 6 years ago
- ☆81Aug 8, 2025Updated last year
- Ultrafast GAN based Vocoder for Text to Speech☆50Jul 16, 2022Updated 4 years ago
- ☆263Dec 8, 2022Updated 3 years ago
- ☆10Sep 3, 2026Updated 2 weeks ago
- Speaker recognition/identification system in Python. Python3 port.☆14May 2, 2015Updated 11 years ago
- File detector, metadata collector and well-formedness checker tool☆18Sep 9, 2026Updated 2 weeks ago