Converts spoken words into text form.
☆78Sep 17, 2025Updated 11 months ago
Alternatives and similar repositories for MAX-Speech-to-Text-Converter
Users that are interested in MAX-Speech-to-Text-Converter are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Generate English-language text similar to the text in the Yelp® review data set.☆18Sep 17, 2025Updated 11 months ago
- Generate English-language text similar to the news articles in the One Billion Words data set.☆26Sep 17, 2025Updated 11 months ago
- PyTorch speech2text inference script for the NVidia openseq2seq wav2letter model variant☆10Aug 12, 2019Updated 7 years ago
- Train a neural network component that can add spatial transformations such as translation and rotation to larger models.☆10Apr 18, 2019Updated 7 years ago
- Python package that can be installed to make it easier to create MAX models☆27May 10, 2021Updated 5 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Identify objects in images using a first-generation deep residual network.☆15Sep 17, 2025Updated 11 months ago
- Identify objects in an image, additionally assigning each pixel of the image to a particular object☆31Sep 17, 2025Updated 11 months ago
- Protect communications with adversarial neural cryptography.☆11Oct 31, 2018Updated 7 years ago
- Create a custom Watson Speech to Text model using specialized domain data☆61Aug 31, 2021Updated 4 years ago
- Generate a new image that mixes the content of a source image with the style of another image.☆50Sep 17, 2025Updated 11 months ago
- Generate personalized recommendations☆14Sep 17, 2025Updated 11 months ago
- 🎧 Automatic Speech Recognition: DeepSpeech & Seq2Seq (TensorFlow)☆222Jun 15, 2020Updated 6 years ago
- Experiment with "one-shot learning" techniques to recognize a voice signature☆24Mar 29, 2020Updated 6 years ago
- Implementation of different noise embeddings for noise aware training of Kaldi acoustic models.☆13Feb 13, 2021Updated 5 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A library of speech gadgets.☆15Oct 15, 2022Updated 3 years ago
- Running Mozilla's implementation of Baidu DeepSpeech on Google Colaboratory☆16Mar 18, 2019Updated 7 years ago
- Answer questions on a given corpus of text.☆32Sep 17, 2025Updated 11 months ago
- This repo contains the baseline model recipes and pre-trained model for GramVanni hindi ASR challenge☆16Mar 26, 2022Updated 4 years ago
- Identify objects in images using a third-generation deep residual network.☆27Sep 17, 2025Updated 11 months ago
- IBM Code Model Asset Exchange: Show and Tell Image Caption Generator☆83Sep 17, 2025Updated 11 months ago
- A simple version of the MAX Object Detector Web App rewritten in python for use in the MAX tutorial☆10Mar 31, 2021Updated 5 years ago
- Generate short audio clips of speech commands and lo-fi instrumental samples☆22Sep 17, 2025Updated 11 months ago
- Identify sounds in short audio clips☆158Sep 17, 2025Updated 11 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Code and videos accompanying the paper "Flickering Adversarial Attacks against Video Recognition Networks"☆16Dec 8, 2022Updated 3 years ago
- Real-time speech enhancement based on spectral subtraction☆16Feb 18, 2018Updated 8 years ago
- Service to evaluate quality measure and cohort specifications against a target patient data set.☆11Jun 2, 2022Updated 4 years ago
- Android Push notifications SDK for IBM Cloud Mobile Services☆10Apr 30, 2021Updated 5 years ago
- This repository provides data and code for "Vox Populi, Vox DIY: Benchmark Dataset for Crowdsourced Audio Transcription" paper.☆16Jul 22, 2021Updated 5 years ago
- From a large speech audio file and its corresponding body of text, automatically chunk the audio and text into (phrase, audio_snippet) pa…☆17May 15, 2015Updated 11 years ago
- Adapt Kaldi-ASR nnet3 chain models from Zamia-Speech.org to a different language model☆33Jan 26, 2020Updated 6 years ago
- Losses and decoders for end-to-end ASR and OCR☆34Oct 30, 2020Updated 5 years ago
- Adversarial Attacks☆21Oct 11, 2021Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Generate embedding vectors from audio files☆58Sep 17, 2025Updated 11 months ago
- DNN-based speech enhancement using Tensorflow by Haoyu Li (Tokyo univ.)☆17Aug 31, 2017Updated 8 years ago
- Detect emotion from audio☆14Nov 20, 2018Updated 7 years ago
- Vim Speech Recognition Experiments☆20May 30, 2025Updated last year
- Evaluation of STT models for german language☆16Jan 22, 2022Updated 4 years ago
- Simple Kaldi model server for chain (nnet3) models in online recognition mode directly from a local microphone☆35Feb 18, 2022Updated 4 years ago
- Adds color to black and white images.☆26Sep 17, 2025Updated 11 months ago