Pytorch based phoneme recognition (TIMIT phoneme classification)
☆35Apr 25, 2018Updated 8 years ago
Alternatives and similar repositories for PytorchSR
Users that are interested in PytorchSR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Phoneme Recognition using RecNet☆97Nov 22, 2016Updated 9 years ago
- Bidirectional dynamic RNN + CTC for phoneme recognition☆47Jun 24, 2020Updated 6 years ago
- Tensorflow implementation of VQVAE for voice conversion☆12Apr 3, 2018Updated 8 years ago
- VQVAE for Unsupervised Voice Conversion☆21Apr 25, 2019Updated 7 years ago
- Build an attention-based model for speech recogntion.Use the Word2vec model to help to train the attention model.☆30Dec 18, 2019Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Network specification and demo☆35Jun 5, 2017Updated 9 years ago
- Phoneme prediction from speech mel-spectrograms using RNN.☆15Jun 4, 2019Updated 7 years ago
- [ICCV'21] The Right to Talk: An Audio-Visual Transformer Approach☆20Aug 2, 2021Updated 5 years ago
- 다양한 feature와 deep learning을 이용한 Phoneme Recognition입니다.☆13Nov 27, 2019Updated 6 years ago
- Voice conversion (VC) investigation using three variants of VAE☆59Oct 28, 2019Updated 6 years ago
- Python wrapper for Sinsy☆53Oct 9, 2023Updated 2 years ago
- MMM 2021: Crossed-Time Delay Neural Network for Speaker Recognition☆11Dec 4, 2021Updated 4 years ago
- This is a working example of using CTC for phone recognition on TIMIT☆50Oct 19, 2017Updated 8 years ago
- Character level speech recognizer using ctc loss with deep rnns in TensorFlow.☆78Jun 9, 2018Updated 8 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Small-footprint Keyword Spotting☆18Jul 28, 2019Updated 7 years ago
- For IEEE ASRU(2025)☆15Jun 21, 2025Updated last year
- Haskell phonology library.☆10Jan 23, 2012Updated 14 years ago
- Alignment examples for Interspeech 2024☆28Jul 5, 2024Updated 2 years ago
- SWIG bindings for Kaldi I/O, built with Conda☆15Dec 15, 2024Updated last year
- A library of speech gadgets.☆15Oct 15, 2022Updated 3 years ago
- Quasi-Periodic Parallel WaveGAN Pytorch implementation☆46Oct 29, 2022Updated 3 years ago
- using world vocoder to extract features and make data for training neural networks☆11Oct 9, 2017Updated 8 years ago
- ☆16Apr 4, 2022Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Transfer Learning from Monolingual ASR to Transcription-free Cross-lingual Voice Conversion☆40Oct 22, 2022Updated 3 years ago
- Spiking neural networks (SNNs) for speech classification☆12Mar 14, 2022Updated 4 years ago
- ☆11Sep 29, 2020Updated 5 years ago
- An unofficial implementation of the paper "One-shot Voice Conversion by Separating Speaker and Content Representations with Instance Norm…☆119May 27, 2021Updated 5 years ago
- PyTorch implementation of Retriever: Learning Content-Style Representation☆12Jan 27, 2023Updated 3 years ago
- ☆20Jun 5, 2022Updated 4 years ago
- kubectl-cred is a kubernetes plugin to which switch to contexts and namespaces and clusters using an interactive CLI.☆15Feb 28, 2022Updated 4 years ago
- Audio Keyword Search☆12May 5, 2019Updated 7 years ago
- USC CS621 Course Project☆26Apr 22, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- BEAST analysis backing effective linguistics☆12Jul 6, 2026Updated last month
- Convolutional Spiking Neural Network to recognize speech utterances using Spike-Timing-Dependent Plasticity☆10Mar 9, 2021Updated 5 years ago
- 以音素建模构建NN-CTC声学模型☆16May 14, 2019Updated 7 years ago
- ☆19Dec 8, 2020Updated 5 years ago
- SyPhon: Constraint-based Learning of Phonological Rules☆11Mar 5, 2025Updated last year
- Pytorch implementation of "Group Latent Embedding for Vector Quantized Variational Autoencoder in Non-Parallel Voice Conversion" [Intersp…☆28Sep 17, 2019Updated 6 years ago
- A pipeline from Dataset Gathering,Data annotations, Model training,Model Evaluation for viseme (visual sound phoneme) classification☆15Jan 19, 2021Updated 5 years ago