Code associated with the paper: Neural Representations for Modeling Variation in Speech.
☆18Mar 10, 2022Updated 4 years ago
Alternatives and similar repositories for neural-acoustic-distance
Users that are interested in neural-acoustic-distance are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Acoustic distance measure for comparing pronunciations☆17Aug 2, 2022Updated 4 years ago
- A pipeline to isolate and transcribe one language in mixed-language speech☆20Oct 25, 2022Updated 3 years ago
- Code associated with the paper: CTC-DRO: Robust Optimization for Reducing Language Disparities in Speech Recognition.☆17May 16, 2025Updated last year
- follow NVIDIA, simplify it and support data parallel.☆13Sep 26, 2019Updated 7 years ago
- Supplementary materials for "Evaluating generalised additive mixed modelling strategies for dynamic speech analysis"☆10Jan 25, 2021Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Web scraper for BibleGateway that will retrieve the entire Bible in a translation of the user's choice to be stored in plain text.☆10Jul 21, 2019Updated 7 years ago
- Self-Supervised Speech Pre-training and Representation Learning Toolkit.☆10Feb 29, 2024Updated 2 years ago
- A module for normalising text.☆10Nov 6, 2019Updated 6 years ago
- MUSDB25 - A Fully Multitrack Dataset for Music Source Separation☆13Mar 29, 2025Updated last year
- NISQA - Non-Intrusive Speech Quality and TTS Naturalness Assessment☆16Apr 13, 2022Updated 4 years ago
- This is a subset of the DALI set consisting of 240 polyphonic recordings that is used to benchmark lyrics transcription evaluation.☆12Nov 30, 2021Updated 4 years ago
- DALI datasets split used to train models presented in the paper Multilingual lyrics-to-audio alignment (ISMIR 2020).☆13May 25, 2021Updated 5 years ago
- This is the accompanying repository to the paper - Automatic Estimation of Singing Voice Musical Dynamics☆16Oct 28, 2024Updated last year
- PyTorch based speaker embedding model☆16Apr 13, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Syllable Segmentation and Cross-Lingual Generalization in a Visually Grounded, Self-Supervised Speech Model☆35Aug 27, 2023Updated 3 years ago
- Tools and scripts for working with ELAN☆10Aug 4, 2022Updated 4 years ago
- ☆28Mar 10, 2017Updated 9 years ago
- This repository shows various ways of deploying a vision model (TensorFlow) from 🤗 Transformers.☆30Aug 22, 2022Updated 4 years ago
- python code for converting among IPA, ARPABET, XSAMPA, Callhome, DISC, TIMIT, plus some lexical tones.☆46Jun 18, 2026Updated 3 months ago
- Code for our CICAI 2022 paper "3D Face Cartoonizer: Generating Personalized 3D Cartoon Faces from 2D Real Photos with a Hybrid Dataset".☆10Aug 9, 2022Updated 4 years ago
- ☆16Jun 13, 2024Updated 2 years ago
- ☆15Nov 26, 2024Updated last year
- Leveraging BERT to Improve Spoken Language Identification☆17Nov 22, 2022Updated 3 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Hearing loss simulation VST plugin☆14Mar 14, 2025Updated last year
- Formant Tracking & Estimation☆87Dec 15, 2024Updated last year
- RStudio Addin that collected several miscellaneous helper functions to save some time☆20Jan 14, 2025Updated last year
- finetune the chain model based on cvte open source model without traing any GMM for frame alignment☆12Aug 6, 2020Updated 6 years ago
- ☆16Dec 6, 2023Updated 2 years ago
- SWIPE Algorithm implementation in Python☆11Dec 24, 2016Updated 9 years ago
- [ISMIR 2022] Transfer Learning of wav2vec 2.0 for Automatic Lyric Transcription☆51May 7, 2024Updated 2 years ago
- Code for "Phoneme Segmentation Using Self-Supervised Speech Models", Strgar & Harwath, Proceedings of the IEEE Spoken Language Technology…☆55Nov 4, 2022Updated 3 years ago
- Transfer learning approach to pronunciation scoring☆12Jan 17, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Overview of corpora/datasets for Germanic low-resource languages and dialects. Accompanies "A Survey of Corpora for Germanic Low-Resource…☆28Feb 16, 2026Updated 7 months ago
- Command line tool for forced-alignment of Spanish speech data☆13Dec 31, 2025Updated 9 months ago
- Charsiu: A neural phonetic aligner.☆351Sep 19, 2022Updated 4 years ago
- ESLTTS dataset☆16Feb 6, 2025Updated last year
- NAISR: A 3D Neural Additive Model for Interpretable Shape Representation☆18Apr 29, 2024Updated 2 years ago
- Code to accompany "Speech denoising with auditory models" (Saddler* & Francl* et al. 2021, Interspeech)☆15Jan 4, 2022Updated 4 years ago
- Tutorial on {Deep} Phonetic Tools given in BigPhon @ LabPhon15☆12Apr 17, 2017Updated 9 years ago