Wave2vec 2.0 Recognize pipeline
☆33Dec 22, 2020Updated 5 years ago
Alternatives and similar repositories for wave2vec-recognize-docker
Users that are interested in wave2vec-recognize-docker are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Pytorch implementation of "Natural TTS Synthesis by Conditioning WaveNet on Mel Spectrogram Predictions", ICASSP, 2018.☆19Jan 21, 2021Updated 5 years ago
- CTC Decoder implementation with python only. Also supports language model decoding using KenLM.☆37May 3, 2024Updated 2 years ago
- Repository for speech paper reading☆33Aug 19, 2021Updated 5 years ago
- Incorporating KenLM language model with HuggingFace implementation of Wav2Vec2CTC Model using beam search decoding☆74Oct 11, 2021Updated 4 years ago
- Review of papers I read☆14Dec 11, 2020Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Python implementation of a simple neural network, including AND, OR, and XOR demos.☆11Jun 13, 2019Updated 7 years ago
- ☆37Jun 9, 2026Updated 2 months ago
- Simple Python library, distributed via binary wheels with few direct dependencies, for easily using wav2vec 2.0 models for speech recogni…☆23Aug 16, 2021Updated 5 years ago
- Dedicated area for the development of data insight in the area of predictive analytics from churn prediction, fraud detection, and gener…☆22Apr 5, 2021Updated 5 years ago
- Small repo describing how to use Hugging Face's Wav2Vec2 with PyCTCDecode☆110Aug 31, 2022Updated 4 years ago
- ☆13Nov 26, 2019Updated 6 years ago
- A simple and humble image captioning application, based on a neural network built with Keras☆10Sep 23, 2022Updated 3 years ago
- Modular and extensible speech recognition library leveraging pytorch-lightning and hydra.☆50May 19, 2021Updated 5 years ago
- 2019 Clova AI Hackathon : Speech - Rank 12 / Team Kai.Lib☆22Jun 11, 2020Updated 6 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- An easy way to fine-tune Wav2Vec 2.0 for low-resource languages.☆81May 20, 2023Updated 3 years ago
- PyTorch implementation of the RNN-based sequence-to-sequence architecture.☆24Jan 21, 2021Updated 5 years ago
- Adnabod lleferydd Cymraeg i'r Gymraeg gyda HuggingFace // Speech Recognition for Welsh with HuggingFace☆13Nov 29, 2022Updated 3 years ago
- ☆22Jul 3, 2019Updated 7 years ago
- Script to generate VAD dataset used in Asteroid recipe☆21Sep 30, 2021Updated 4 years ago
- ASR project with pytorch-lightning☆20Mar 21, 2025Updated last year
- Hosts text-to-speech corpus and speech synthesizers for African languages.☆20May 31, 2023Updated 3 years ago
- ☆13Aug 7, 2021Updated 5 years ago
- Pytorch implementation of Deepmind's WaveRNN model☆13Apr 5, 2020Updated 6 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Counterfactual SHAP: a framework for counterfactual feature importance☆21Jul 6, 2023Updated 3 years ago
- A PyTorch Implementation of "Attention Is All You Need"☆38Oct 3, 2021Updated 4 years ago
- ☆15Sep 26, 2022Updated 3 years ago
- Text Classification model deployment using FastAPI, Streamlit and Docker Compose☆15Feb 12, 2021Updated 5 years ago
- The codebase for Data-driven general-purpose voice activity detection.☆93Aug 3, 2023Updated 3 years ago
- Social previews generator as a microservice.☆12Apr 9, 2022Updated 4 years ago
- Acoustic distance measure for comparing pronunciations☆17Aug 2, 2022Updated 4 years ago
- Experiments with Hugging Face 🔬 🤗☆48Apr 18, 2026Updated 4 months ago
- Korean speech recognition based on transformer (트랜스포머 기반 한국어 음성 인식)☆31Feb 19, 2021Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A Python Implementation of Driedger's "Let It Bee" Technique for Audio Mosaicing☆25Sep 14, 2024Updated last year
- Neuro-Holistic Audio-eNhancement System (N-HANS)☆40Jun 4, 2021Updated 5 years ago
- Give me a term and I'll give you a list of links found in its Wikipedia article☆15Nov 30, 2017Updated 8 years ago
- Use quantized versions of Whisper to speed up inference☆12Oct 16, 2024Updated last year
- Transcribing audio files using Hugging Face's implementation of Wav2Vec2 + "chain-linking" NLP tasks to combine speech-to-text with downs…☆32Mar 20, 2021Updated 5 years ago
- ☆13Dec 3, 2019Updated 6 years ago
- Transformer implementation speciaized in speech recognition tasks using Pytorch.☆65Nov 28, 2021Updated 4 years ago