SailAlign is an open-source software toolkit for robust long speech-text alignment implementing an adaptive, iterative speech recognition and text alignment scheme that allows for the processing of very long (and possibly noisy) audio and is robust to transcription errors. It is mainly written as a perl library but its functionality also depends…
☆99Apr 5, 2022Updated 4 years ago
Alternatives and similar repositories for sail_align
Users that are interested in sail_align are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Python interface for forced audio alignment using HTK and SoX☆351Jun 28, 2020Updated 6 years ago
- Deploy Kaldi models using grpc for bidirectional streaming.☆17Sep 30, 2024Updated 2 years ago
- Script for converting kaldi GMM/HMM models to HTK format☆11Jul 18, 2024Updated 2 years ago
- Long audio alignment using Kaldi☆23Apr 22, 2021Updated 5 years ago
- Read and write HTK and HTS files from python.☆20Mar 17, 2015Updated 11 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Utils and modules for Speech Language and Multimodal processing using pytorch and pytorch lightning☆22Feb 16, 2023Updated 3 years ago
- gentle forced aligner☆1,711Jul 24, 2026Updated 2 months ago
- python wrap for hts engine☆14Jan 30, 2018Updated 8 years ago
- Training and using classifiers for textual documents☆15Sep 16, 2016Updated 10 years ago
- NMT based punctuation prediction system using lexical and acoustic features .☆14Mar 30, 2020Updated 6 years ago
- A ROS framework for Audio Analysis☆12Apr 5, 2017Updated 9 years ago
- A collection of links and notes on forced alignment tools☆943Jul 22, 2026Updated 2 months ago
- Python implementation of CTC beam search decoder + agnostic LM scorer☆20Dec 16, 2020Updated 5 years ago
- C++ implementation of End to End TTS which combines both Tacatron2 and LPCNET Vocoder.☆32Oct 1, 2019Updated 7 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- PyTorch speech2text inference script for the NVidia openseq2seq wav2letter model variant☆10Aug 12, 2019Updated 7 years ago
- A script for audio/transcript alignment. Fork of p2fa.☆70Mar 15, 2018Updated 8 years ago
- A collection of utilities for handling IPA phones.☆27Sep 24, 2023Updated 3 years ago
- ☆17Apr 8, 2016Updated 10 years ago
- Mining effective negative training samples for keyword spotting (PyTorch)☆68May 23, 2020Updated 6 years ago
- ☆13Jun 30, 2026Updated 3 months ago
- Calculates the Word Error Rate between two text files☆20Nov 10, 2022Updated 3 years ago
- R Code recipes for Functional Data Analysis for phonetic analysis.☆13Jul 31, 2024Updated 2 years ago
- ☆47Oct 24, 2020Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A repository for maintaing the fave-align and fave-extract toolkits☆118Mar 29, 2024Updated 2 years ago
- Supplementary materials for "Evaluating generalised additive mixed modelling strategies for dynamic speech analysis"☆10Jan 25, 2021Updated 5 years ago
- FastCGI support for Kaldi ASR☆185Apr 5, 2019Updated 7 years ago
- Feature set algebra for linguistics☆17Jul 7, 2026Updated 3 months ago
- This code implements a basic MLP for speech recognition. The MLP is trained with pytorch, while feature extraction, alignments, and dec…☆40Feb 10, 2018Updated 8 years ago
- Text-to-Speech tutorial at SLTU 2016☆35May 10, 2016Updated 10 years ago
- JSON schema and JavaScript model classes for dealing with time-aligned transcripts of speech.☆16Aug 20, 2018Updated 8 years ago
- ☆24Sep 25, 2018Updated 8 years ago
- An LSTM RNN for restoring missing punctuation in unsegmented text.☆78Sep 24, 2016Updated 10 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Converts an audio file to a 3D spectrogram and (optionally) saves as a stereolithography (STL) file for 3D printing☆22Oct 31, 2021Updated 4 years ago
- how to generate the full-contextual labels from un-seen text for the application of HMM-based speech synthesis (HTS)☆12Nov 22, 2019Updated 6 years ago
- GSoC'16 RedHen Labs☆11Aug 22, 2016Updated 10 years ago
- Tool for creating Kaldi nnet3 recipes using the International Phonetic Alphabet (IPA)☆10Jun 2, 2021Updated 5 years ago
- Aligns text (lyrics) with monophonic singing voice (audio). The algorithm uses structural segmentation to segment the audio into structur…☆94Feb 13, 2018Updated 8 years ago
- TranscriberAG open development☆40Nov 16, 2014Updated 11 years ago
- Wrapper to pocketsphinx phoneme labeling tools☆18Sep 9, 2016Updated 10 years ago