ForceAlign is a Python library for forced alignment of English text to English audio. You can use ForceAlign to get word or phoneme level text alignments of audio, with each word or phoneme's start and end time within the audio. ForceAlign was designed to be easy to install and use, without requiring any third-party, non-Python dependencies.
☆28Dec 4, 2024Updated last year
Alternatives and similar repositories for forcealign
Users that are interested in forcealign are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Cochlear implant signal processing☆10Jun 24, 2021Updated 5 years ago
- faster inference☆27Jan 20, 2025Updated last year
- Decoders from Kaldi using OpenFst☆35Apr 10, 2026Updated 4 months ago
- ☆29Aug 8, 2024Updated 2 years ago
- Package for easy handle mobi books in swift☆13Feb 5, 2026Updated 6 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Zeta implementation of a reusable and plug in and play feedforward from the paper "Exponentially Faster Language Modeling"☆16Nov 11, 2024Updated last year
- Transfer learning approach to pronunciation scoring☆12Jan 17, 2024Updated 2 years ago
- MagicData-RAMC Dataset and Baseline☆64Sep 13, 2022Updated 3 years ago
- This repository is a repository for the paper, "Irgun: Improved residue based gradual up-scaling network for single image super resolutio…☆16Aug 26, 2020Updated 6 years ago
- a compact audio-to-phoneme aligner for singing voice☆12Jan 17, 2024Updated 2 years ago
- ✒️ LanguageTool integration for Quill.js editors☆17Aug 20, 2024Updated 2 years ago
- Implementation of the LDP module block in PyTorch and Zeta from the paper: "MobileVLM: A Fast, Strong and Open Vision Language Assistant …☆15Mar 11, 2024Updated 2 years ago
- [INTERSPEECH 2023] Knowledge Transfer from Pre-trained Language Models to Cif-based Recognizers via Hierarchical Distillation☆41Jul 14, 2026Updated last month
- This repo contains script to download MUSIC dataset from youtube☆12Jan 19, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Supplementary materials for "Evaluating generalised additive mixed modelling strategies for dynamic speech analysis"☆10Jan 25, 2021Updated 5 years ago
- Talk To AI with FastRTC enables natural, real-time voice conversations with AI using WebRTC, offering customizable voices, interfaces, an…☆47Mar 10, 2025Updated last year
- ☆13Aug 12, 2026Updated 2 weeks ago
- Convert various blog dumps to a standard JSON☆12Mar 29, 2026Updated 5 months ago
- Some stuff to handle various datasets☆15Mar 2, 2018Updated 8 years ago
- ☆130Aug 3, 2026Updated 3 weeks ago
- A command-line interface for running Supertonic TTS models using MNN.☆17Jun 22, 2026Updated 2 months ago
- This is a Document to Handwriting a website using HTML, CSS, JS and Google font API. We type our work in the text box and our work will b…☆21Feb 14, 2023Updated 3 years ago
- Unbounded cache model for online language modeling with open vocabulary☆11Feb 15, 2019Updated 7 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Streamable Text-to-Speech model using a language modeling approach, without vector quantization☆109May 20, 2025Updated last year
- A lightweight tool that efficiently isolates target speaker data from your datasets.☆19Nov 23, 2024Updated last year
- Fast and accurate natural language detection. Detector written in Python. Nito-ELD, ELD.☆22Jul 19, 2026Updated last month
- [ICLR'25] "Understanding Bottlenecks of State Space Models through the Lens of Recency and Over-smoothing" by Peihao Wang, Ruisi Cai, Yue…☆18Mar 21, 2025Updated last year
- A sample Xcode Project to run Python in Xcode.☆13Apr 1, 2022Updated 4 years ago
- audio/speech feature extraction using parselmouth, librosa, disvoice☆10Jan 28, 2022Updated 4 years ago
- Detect frames in a Manga page with OpenCV☆10Nov 19, 2016Updated 9 years ago
- Curriculum Vitae of Quan Wang☆15Updated this week
- audiolm-pytorch training code☆15Jul 31, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Sequence alignement methods with helpers for PyTorch.☆24Nov 30, 2022Updated 3 years ago
- Provide the best of TED.com for offline usage!☆21Aug 4, 2026Updated 3 weeks ago
- Wenet speech to text for react native☆10Nov 1, 2022Updated 3 years ago
- An original package of the dynamic compressive gammachirp filterbank (dcGC-FB)☆14Aug 20, 2026Updated last week
- Tobii Eye Tracker 4C Naïve Solution☆21Feb 25, 2021Updated 5 years ago
- ☆11Sep 30, 2021Updated 4 years ago
- Data manipulation and transformation for audio signal processing, powered by PyTorch☆10Sep 30, 2024Updated last year