Converts spoken words into text form.
☆78Sep 17, 2025Updated 10 months ago
Alternatives and similar repositories for MAX-Speech-to-Text-Converter
Users that are interested in MAX-Speech-to-Text-Converter are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Generate English-language text similar to the text in the Yelp® review data set.☆18Sep 17, 2025Updated 10 months ago
- Generate English-language text similar to the news articles in the One Billion Words data set.☆26Sep 17, 2025Updated 10 months ago
- ☆17Jul 15, 2026Updated last week
- Generate a summarized description of a body of text☆27Sep 17, 2025Updated 10 months ago
- PyTorch speech2text inference script for the NVidia openseq2seq wav2letter model variant☆10Aug 12, 2019Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Train a neural network component that can add spatial transformations such as translation and rotation to larger models.☆10Apr 18, 2019Updated 7 years ago
- Python package that can be installed to make it easier to create MAX models☆27May 10, 2021Updated 5 years ago
- Identify objects in images using a first-generation deep residual network.☆15Sep 17, 2025Updated 10 months ago
- Create a custom Watson Speech to Text model using specialized domain data☆61Aug 31, 2021Updated 4 years ago
- Generate a new image that mixes the content of a source image with the style of another image.☆50Sep 17, 2025Updated 10 months ago
- 🎧 Automatic Speech Recognition: DeepSpeech & Seq2Seq (TensorFlow)☆222Jun 15, 2020Updated 6 years ago
- Experiment with "one-shot learning" techniques to recognize a voice signature☆24Mar 29, 2020Updated 6 years ago
- Implementation of different noise embeddings for noise aware training of Kaldi acoustic models.☆13Feb 13, 2021Updated 5 years ago
- A library of speech gadgets.☆15Oct 15, 2022Updated 3 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Running Mozilla's implementation of Baidu DeepSpeech on Google Colaboratory☆16Mar 18, 2019Updated 7 years ago
- Answer questions on a given corpus of text.☆32Sep 17, 2025Updated 10 months ago
- This repo contains the baseline model recipes and pre-trained model for GramVanni hindi ASR challenge☆16Mar 26, 2022Updated 4 years ago
- Generate short audio clips of speech commands and lo-fi instrumental samples☆22Sep 17, 2025Updated 10 months ago
- Identify sounds in short audio clips☆158Sep 17, 2025Updated 10 months ago
- Code and videos accompanying the paper "Flickering Adversarial Attacks against Video Recognition Networks"☆16Dec 8, 2022Updated 3 years ago
- Service to evaluate quality measure and cohort specifications against a target patient data set.☆11Jun 2, 2022Updated 4 years ago
- Experiments with Hugging Face 🔬 🤗☆47Apr 18, 2026Updated 3 months ago
- This repository provides data and code for "Vox Populi, Vox DIY: Benchmark Dataset for Crowdsourced Audio Transcription" paper.☆16Jul 22, 2021Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- From a large speech audio file and its corresponding body of text, automatically chunk the audio and text into (phrase, audio_snippet) pa…☆17May 15, 2015Updated 11 years ago
- Locate and tag named entities in text☆25Sep 17, 2025Updated 10 months ago
- Losses and decoders for end-to-end ASR and OCR☆34Oct 30, 2020Updated 5 years ago
- Adversarial Attacks☆21Oct 11, 2021Updated 4 years ago
- DNN-based speech enhancement using Tensorflow by Haoyu Li (Tokyo univ.)☆17Aug 31, 2017Updated 8 years ago
- MAX Optical Character Recognition☆51Sep 17, 2025Updated 10 months ago
- Vim Speech Recognition Experiments☆20May 30, 2025Updated last year
- Evaluation of STT models for german language☆16Jan 22, 2022Updated 4 years ago
- Simple Kaldi model server for chain (nnet3) models in online recognition mode directly from a local microphone☆35Feb 18, 2022Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- An application that measures the respiratory rate for patients using a flir one mobile thermal camera☆12Aug 5, 2020Updated 5 years ago
- 📖 LanMIT: A Toolkit for Improving Language Models in Low-resourced Speech Recognition based on Kaldi.☆22Jul 12, 2019Updated 7 years ago
- wake word spotting with kaldi☆19Dec 3, 2020Updated 5 years ago
- Train a Deep Learning model to classify audio embeddings on IBM's Deep Learning as a Service (DLaaS) platform - Watson Machine Learning☆102Sep 17, 2025Updated 10 months ago
- Korean ASR Corpus generated from TEDx talks☆27Jan 11, 2019Updated 7 years ago
- ☆17Nov 25, 2019Updated 6 years ago
- Implementation of Z-BERT-A: a zero-shot pipeline for unknown intent detection.☆44Jun 13, 2023Updated 3 years ago