Simple automatic speech recognition system based on digits corpora (Polish language), created in Kaldi toolkit. Despite of the language difference, this is an effect of 'Kaldi for dummies' tutorial published in kaldi-help discussion group. No audio data - this is just an example.
☆11May 29, 2016Updated 10 years ago
Alternatives and similar repositories for kaldifordummies
Users that are interested in kaldifordummies are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Neural Network Semantic Parser for Almond☆15Apr 11, 2019Updated 7 years ago
- Deep Learning For Ultrasound Tongue Imaging☆13Dec 17, 2024Updated last year
- A step-by-step problem set for implementing a high-quality deep dependency parser in Pytorch☆15Aug 12, 2017Updated 9 years ago
- ctc_beamsearch☆18Oct 26, 2016Updated 9 years ago
- PREDICTING TONGUE MOTION IN UNLABELED ULTRASOUND VIDEOS USING CONVOLUTIONAL LSTM NEURAL NETWORKS☆19Oct 29, 2018Updated 7 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Deep understanding and modelling of the hierarchical structure of prosody☆25May 12, 2019Updated 7 years ago
- Code for TALLIP2019 paper "µ-Forcing: Training Variational Recurrent Autoencoders for Text Generation"☆12May 27, 2019Updated 7 years ago
- Neural Language Models as Psycholinguistic Subjects: Representations of Syntactic State☆17Mar 4, 2019Updated 7 years ago
- A series of Jupyter notebooks on signal processing☆53Dec 16, 2018Updated 7 years ago
- A PyTorch implementation of DeepSpeech and DeepSpeech2.☆50Dec 4, 2018Updated 7 years ago
- Code for paper titled "Using generative modelling to produce varied intonation for speech synthesis" submitted to the Speech Synthesis Wo…☆24Dec 8, 2019Updated 6 years ago
- MTracker is a tool for automatic splining tongue shapes in ultrasound images by harnessing the power of deep convolutional neural network…☆20Feb 12, 2021Updated 5 years ago
- Data processing tools for preparing speech and labels for training TTS voices☆29Aug 13, 2020Updated 6 years ago
- Code accompanying the paper "Effective Estimation of Deep Generative Language Models".☆24May 1, 2020Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- An Ellipsis-aware Chinese Dependency Treebank for Web Text☆27May 14, 2018Updated 8 years ago
- A repository for the EMNLP 2021 paper "Is Information Density Uniform in Task-Oriented Dialogues?" and for the CoNLL 2021 paper "Analysin…☆10Jun 17, 2024Updated 2 years ago
- Text-Independent Speaker Recognition Using Gaussian Mixture Models☆12Jul 1, 2015Updated 11 years ago
- A public dataset containing chord/beat annotation from a music game named 'osu!'.☆13Oct 17, 2017Updated 8 years ago
- ABX discrimination task in python☆45Oct 7, 2024Updated last year
- 2D U-Net using deformable convolution☆28Dec 12, 2020Updated 5 years ago
- MATLAB model of the auditory periphery☆17Nov 28, 2011Updated 14 years ago
- Train a LSTM neural networks on Vox Forge public audio data set to recognize speaker's gender☆13Mar 26, 2026Updated 5 months ago
- Emacs Board of SMTH☆18Apr 12, 2013Updated 13 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Representations of language in a model of visually grounded speech signal.☆23Apr 19, 2018Updated 8 years ago
- Just a mirror, this is not the official repository of emacs-w3m.☆20Mar 28, 2012Updated 14 years ago
- Java API for the online speech recognition services provided by phon.ioc.ee☆20Jun 4, 2021Updated 5 years ago
- Use it to convert a whole directory from Python 2 to Python 3, including IPython Notebooks☆10Nov 23, 2015Updated 10 years ago
- ☆14Dec 7, 2018Updated 7 years ago
- Morphological analysis and generation of Amharic, Oromo, and Tigrinya☆13Feb 18, 2017Updated 9 years ago
- Unsupervised Learning for Optical Flow Estimation Using Pyramid Convolution LSTM.☆35Jul 29, 2019Updated 7 years ago
- A neural language model that estimates incremental processing complexity☆41Oct 27, 2021Updated 4 years ago
- VoxLingua107 recipe for SpeechBrain☆13Jul 3, 2021Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Mastodon server running for the Doubanius Tertius project☆10Apr 4, 2022Updated 4 years ago
- ☆51Feb 15, 2019Updated 7 years ago
- Code for replicating the work in "Targeted Syntactic Evaluation of Language Models." EMNLP 2018.☆44Apr 25, 2020Updated 6 years ago
- Analysis and investigating the confounding effect of accents in end-to-end Automatic Speech Recognition models.☆15Jun 27, 2020Updated 6 years ago
- Code for the paper LazImpa: Lazy and Impatient neural agents learn to communicate efficiently. Mathieu Rita, Rahma Chaabouni and Emmanuel…☆17Nov 21, 2020Updated 5 years ago
- Emergent Communication of Generalizations, NeurIPS 2021☆13Sep 29, 2021Updated 4 years ago
- Code and Results for "Universals of word order reflect optimization of grammars for efficient communication"☆14Aug 5, 2022Updated 4 years ago