Speech formant tracking code in Python
☆15Oct 10, 2013Updated 12 years ago
Alternatives and similar repositories for fun-with-formants
Users that are interested in fun-with-formants are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of "Improving Whispered Speech Recognition Performance using Pseudo-whispered based Data Augmentation"☆14Oct 31, 2024Updated last year
- Formant Tracking & Estimation☆83Dec 15, 2024Updated last year
- Synthesis of percussion sounds using sinusoidal modelling, DDSP noise synthesis, and a neural source filter approach.☆35Jan 7, 2025Updated last year
- Phonetic Analysis ToolKIT - PATKIT - Python package for analysing phonetic data☆11Jun 27, 2026Updated last month
- Official source code for the paper "Tailored Design of Audio-Visual Speech Recognition Models using Branchformers"☆15Feb 24, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- software for editing dynamic formant measurements☆15Aug 16, 2015Updated 10 years ago
- Interspeech Tutorial - Resource Efficient and Cross-Modal Learning Toward Foundation Modeling☆15Oct 9, 2023Updated 2 years ago
- Praat syntax highlighter for Sublime Text☆19May 22, 2015Updated 11 years ago
- This package contains functions for converting wav files into auditory representations and comparing them☆56May 19, 2025Updated last year
- (SLT 2024) Learning Video Temporal Dynamics with Cross-Modal Attention for Robust Audio-Visual Speech Recognition☆13Oct 22, 2024Updated last year
- TheGlueNote is representation model for note-wise music alignment.☆14Jul 19, 2024Updated 2 years ago
- An implement of SPEECHSPLIT☆15Sep 12, 2020Updated 5 years ago
- explore AMT from the perspective of timbre☆26Apr 17, 2026Updated 3 months ago
- ☆29Jul 12, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Annotations and scripts for use with University of Wisconsin X-Ray Microbeam Speech Production Database (1994)☆14Oct 8, 2020Updated 5 years ago
- ☆17Jan 1, 2024Updated 2 years ago
- ☆25Sep 27, 2022Updated 3 years ago
- Landing Page☆11May 7, 2026Updated 2 months ago
- (ICLR 2025) Multi-Task Corrupted Prediction for Learning Robust Audio-Visual Speech Representation☆16Apr 29, 2025Updated last year
- HRTF data preparation for machine learning by finding common measurement angles☆12May 14, 2019Updated 7 years ago
- Formant Extraction☆15Oct 22, 2013Updated 12 years ago
- a close enough approximation of the shadertoy framework☆12Jul 2, 2020Updated 6 years ago
- A python implementation of a simple Unit Selection Text-to-Speech (TTS) synthesis system. It works with CMU-Arctic data by default☆11Mar 14, 2015Updated 11 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- EMPHASIS: An Emotional Phoneme-based Acoustic Model for Speech Synthesis System☆15Mar 31, 2019Updated 7 years ago
- Mirror of the Auditory Modelling Toolbox http://amtoolbox.sourceforge.net/☆11Jan 28, 2019Updated 7 years ago
- A fasttrack implementation in python☆13Feb 10, 2026Updated 5 months ago
- [SpeechCom Journal] Learning and controlling the source-filter representation of speech with a variational autoencoder☆46Apr 18, 2023Updated 3 years ago
- A Bayesian method which utilises the rich structure embedded in the sensing matrix for fast sparse signal recovery☆11Apr 25, 2018Updated 8 years ago
- ☆13Apr 4, 2023Updated 3 years ago
- Interface for Controllable Expressive Talking Machine☆40Sep 20, 2025Updated 10 months ago
- Single Pass Spectrogram Inversion in a Jupyter Python notebook☆34Aug 10, 2017Updated 8 years ago
- This repository contains laughter-related synthesis systems.☆13Nov 7, 2020Updated 5 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- WildVSR☆22Dec 13, 2023Updated 2 years ago
- Interface for running Praat scripts through Python☆17May 16, 2025Updated last year
- Python API for SOFA (Spatially Oriented Format for Acoustics)☆10Aug 13, 2019Updated 6 years ago
- Script to perform statistical significance test between ASR hypotheses.☆23Aug 13, 2017Updated 8 years ago
- GE2E Speaker Encoder - Generalized End-To-End Loss for Speaker Verification☆14May 17, 2020Updated 6 years ago
- ☆14Sep 27, 2019Updated 6 years ago
- Converts an audio file to a 3D spectrogram and (optionally) saves as a stereolithography (STL) file for 3D printing☆22Oct 31, 2021Updated 4 years ago