Introduction to Speech Processing
☆122May 12, 2026Updated 2 months ago
Alternatives and similar repositories for itsp
Users that are interested in itsp are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- I wanted guided tutorials on digital signal processing so I decided to create them. The result is this ebook: "Digital Signal Processing …☆12Feb 5, 2024Updated 2 years ago
- Self-Supervised Speech Pre-training and Representation Learning Toolkit.☆10Feb 29, 2024Updated 2 years ago
- ☆18Sep 8, 2025Updated 10 months ago
- Sound field reconstruction using neural processes with dynamic kernels☆16Mar 25, 2025Updated last year
- SAMO: SPEAKER ATTRACTOR MULTI-CENTER ONE-CLASS LEARNING FOR VOICE ANTI-SPOOFING☆42Apr 5, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This is a repository of neural full-rank spatial covariance analysis with speaker activity (neural FCASA).☆40Mar 12, 2025Updated last year
- A workflow for acoustic analysis of speech prosody based on continuous measurements of periodic energy and F0 (requires Praat and R). Pro…☆15Apr 29, 2026Updated 2 months ago
- Voice Framework☆18Jan 21, 2026Updated 6 months ago
- Klatt formant synthesizer☆76Jun 26, 2026Updated 3 weeks ago
- The LAP Challenge aims at advancing spatial audio technologies through the personalization of HRTFs.☆16Aug 12, 2025Updated 11 months ago
- ☆13Jan 14, 2025Updated last year
- Measuring impulse response with time-stretched pulse (TSP) signal☆14Jul 3, 2019Updated 7 years ago
- Praat-based tools for spectral analysis☆37May 28, 2026Updated last month
- Neural IIR Filter Field for HRTF Upsampling and Personalization☆29Feb 26, 2024Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆26Apr 30, 2026Updated 2 months ago
- Supplementary materials for "Evaluating generalised additive mixed modelling strategies for dynamic speech analysis"☆10Jan 25, 2021Updated 5 years ago
- Praat-based tools for EGG analysis☆20Sep 21, 2023Updated 2 years ago
- Cover Song Detection System☆10Mar 29, 2019Updated 7 years ago
- Implementation of "Improving Whispered Speech Recognition Performance using Pseudo-whispered based Data Augmentation"☆14Oct 31, 2024Updated last year
- PyTorch-based room impulse response (RIR) simulation toolkit with dynamic scenes, GPU acceleration.☆22Feb 18, 2026Updated 5 months ago
- This repository contains the code for the paper: "DeToxy: A Large-Scale Multimodal Dataset for Toxicity Classification in Spoken Utteranc…☆21Oct 13, 2022Updated 3 years ago
- PyTorch implementation of "Source Separation by Flow Matching (FLOSS)" by Google DeepMind☆96Nov 24, 2025Updated 7 months ago
- Praat script for automatic formant optimization☆15Jan 27, 2023Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code for "Phoneme Segmentation Using Self-Supervised Speech Models", Strgar & Harwath, Proceedings of the IEEE Spoken Language Technology…☆55Nov 4, 2022Updated 3 years ago
- ☆18Apr 24, 2025Updated last year
- Causality Check in Frame-online Speech Separation☆51Dec 11, 2022Updated 3 years ago
- Pitch-shifting, time-stretching, and vocoding of speech with Controllable LPCNet (CLPCNet)☆166Aug 5, 2022Updated 3 years ago
- Trainable algorithm for automatic measurement of voice onset time☆69Jul 26, 2023Updated 2 years ago
- JATTS: A modern, research-oriented Japanese Text-to-speech Open-sourced Toolkit☆43Mar 13, 2026Updated 4 months ago
- A Jupyter book accompanying the ISMIR 2023 tutorial Introduction to DIfferentiable Audio Synthesiser Programming☆62Jun 30, 2025Updated last year
- These are Jupyter Notebooks to help guide people to learn how to use Praat-Parselmouth☆43Sep 29, 2021Updated 4 years ago
- Ablation study of local spectral attention (LSA) for full-band speech enhancement (SE)☆28Sep 16, 2023Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Here is a repository stored the classical sound source localization algorithms in spherical domain, namely, PWD, DAS, SHMUSIC, SHMVDR, S…☆23Nov 16, 2023Updated 2 years ago
- Single channel speech source separation by diffusion process (ICASSP 2023)☆126Mar 15, 2024Updated 2 years ago
- Official Code for SyllableLM: Learning Coarse Semantic Units for Speech Language Models☆63Jul 1, 2025Updated last year
- My attempts at applying Soundstream design on learned tokenization of text and then applying hierarchical attention to text generation☆90Oct 11, 2024Updated last year
- Praat textgrid manipulation in Python☆55Apr 3, 2025Updated last year
- Operator-level compressed GTCRN with ERB-CRM pipeline preserved and DPGRNN intact, ready for edge deployment.☆22Feb 11, 2026Updated 5 months ago
- A Pitch shifter plugin implementation using JUCE and rubberband☆19Jan 29, 2023Updated 3 years ago