SailAlign is an open-source software toolkit for robust long speech-text alignment implementing an adaptive, iterative speech recognition and text alignment scheme that allows for the processing of very long (and possibly noisy) audio and is robust to transcription errors. It is mainly written as a perl library but its functionality also depends…
☆99Apr 5, 2022Updated 4 years ago
Alternatives and similar repositories for sail_align
Users that are interested in sail_align are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Python interface for forced audio alignment using HTK and SoX☆351Jun 28, 2020Updated 6 years ago
- Deploy Kaldi models using grpc for bidirectional streaming.☆17Sep 30, 2024Updated last year
- Script for converting kaldi GMM/HMM models to HTK format☆11Jul 18, 2024Updated 2 years ago
- Long audio alignment using Kaldi☆23Apr 22, 2021Updated 5 years ago
- Read and write HTK and HTS files from python.☆20Mar 17, 2015Updated 11 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Utils and modules for Speech Language and Multimodal processing using pytorch and pytorch lightning☆22Feb 16, 2023Updated 3 years ago
- gentle forced aligner☆1,704Updated this week
- python wrap for hts engine☆14Jan 30, 2018Updated 8 years ago
- Training and using classifiers for textual documents☆15Sep 16, 2016Updated 9 years ago
- NMT based punctuation prediction system using lexical and acoustic features .☆14Mar 30, 2020Updated 6 years ago
- A ROS framework for Audio Analysis☆12Apr 5, 2017Updated 9 years ago
- A collection of links and notes on forced alignment tools☆942Jul 22, 2026Updated last week
- Python implementation of CTC beam search decoder + agnostic LM scorer☆20Dec 16, 2020Updated 5 years ago
- C++ implementation of End to End TTS which combines both Tacatron2 and LPCNET Vocoder.☆32Oct 1, 2019Updated 6 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- PyTorch speech2text inference script for the NVidia openseq2seq wav2letter model variant☆10Aug 12, 2019Updated 6 years ago
- A script for audio/transcript alignment. Fork of p2fa.☆69Mar 15, 2018Updated 8 years ago
- A collection of utilities for handling IPA phones.☆27Sep 24, 2023Updated 2 years ago
- ☆17Apr 8, 2016Updated 10 years ago
- Mining effective negative training samples for keyword spotting (PyTorch)☆66May 23, 2020Updated 6 years ago
- ☆13Jun 30, 2026Updated last month
- Calculates the Word Error Rate between two text files☆20Nov 10, 2022Updated 3 years ago
- R Code recipes for Functional Data Analysis for phonetic analysis.☆13Jul 31, 2024Updated last year
- ☆45Oct 24, 2020Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A repository for maintaing the fave-align and fave-extract toolkits☆118Mar 29, 2024Updated 2 years ago
- Supplementary materials for "Evaluating generalised additive mixed modelling strategies for dynamic speech analysis"☆10Jan 25, 2021Updated 5 years ago
- FastCGI support for Kaldi ASR☆185Apr 5, 2019Updated 7 years ago
- Feature set algebra for linguistics☆17Jul 7, 2026Updated 3 weeks ago
- This code implements a basic MLP for speech recognition. The MLP is trained with pytorch, while feature extraction, alignments, and dec…☆40Feb 10, 2018Updated 8 years ago
- Text-to-Speech tutorial at SLTU 2016☆35May 10, 2016Updated 10 years ago
- JSON schema and JavaScript model classes for dealing with time-aligned transcripts of speech.☆16Aug 20, 2018Updated 7 years ago
- ☆24Sep 25, 2018Updated 7 years ago
- An LSTM RNN for restoring missing punctuation in unsegmented text.☆78Sep 24, 2016Updated 9 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Converts an audio file to a 3D spectrogram and (optionally) saves as a stereolithography (STL) file for 3D printing☆22Oct 31, 2021Updated 4 years ago
- how to generate the full-contextual labels from un-seen text for the application of HMM-based speech synthesis (HTS)☆12Nov 22, 2019Updated 6 years ago
- GSoC'16 RedHen Labs☆11Aug 22, 2016Updated 9 years ago
- Tool for creating Kaldi nnet3 recipes using the International Phonetic Alphabet (IPA)☆10Jun 2, 2021Updated 5 years ago
- Aligns text (lyrics) with monophonic singing voice (audio). The algorithm uses structural segmentation to segment the audio into structur…☆94Feb 13, 2018Updated 8 years ago
- Wrapper to pocketsphinx phoneme labeling tools☆18Sep 9, 2016Updated 9 years ago
- Autoregressive HMM version of the HTS demo for statistical speech synthesis (includes autoregressive clustering)☆16Sep 12, 2014Updated 11 years ago