wav2vec2 asr with transformers
☆16Oct 26, 2021Updated 4 years ago
Alternatives and similar repositories for wav2vec2-asr
Users that are interested in wav2vec2-asr are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Convert any photos or videos to artwork. Implementation of Style Transfer in TensorFlow.☆12Jan 26, 2022Updated 4 years ago
- ☆10Aug 26, 2021Updated 4 years ago
- Fine-tuning Wav2Vec2.0 on Common Voice(zh-HK)☆16May 8, 2022Updated 4 years ago
- ☆11Oct 24, 2022Updated 3 years ago
- ☆37Nov 19, 2020Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Speech understanding system training toolkit, including tasks of ASR, SSL, LM, etc.☆12Feb 12, 2026Updated 5 months ago
- Source code to "SliTraNet: Automatic Detection of Slide Transitions in Lecture Videos using Convolutional Neural Networks"☆10Dec 17, 2023Updated 2 years ago
- An simple pytorch implementation of Flash MultiHead Attention☆22Feb 5, 2024Updated 2 years ago
- Source code for MWO2KG and Echidna: Constructing and Exploring Knowledge Graphs from Maintenance Data☆10Feb 13, 2023Updated 3 years ago
- Final training script from HuggingFace Whisper Fine tuning event - to get best results on finetuned model.☆12Dec 24, 2022Updated 3 years ago
- My graduation project.☆13Oct 12, 2023Updated 2 years ago
- Decoding of LDPC Codes Using the Information Bottleneck Method in Python☆17Dec 11, 2018Updated 7 years ago
- A class that is able to develop any (n, k) linear code. Includes an implementation of an ASCII correcting linear code along with a simula…☆12Apr 20, 2025Updated last year
- This project focuses on the classification of animal sounds using deep learning. The core idea is to utilize audio processing techniques …☆10Dec 3, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Framework to train CNN and use it in Relaxed Projected Gradient Descent (RPGD) to reconstruct images☆13Nov 25, 2019Updated 6 years ago
- Tools for Natural Language Text aware PDF structure analysis☆15Mar 11, 2022Updated 4 years ago
- Reinforcement Learning for Bit Flipping decoding of linear codes☆14Sep 12, 2020Updated 5 years ago
- fine-tune Wav2vec2. an ASR model released by Facebook☆36Dec 11, 2021Updated 4 years ago
- Hierarchical Context Tagger for utterance rewriting☆13Mar 27, 2022Updated 4 years ago
- This repository contains the VLEngagement dataset and the helper functions/ tools required to work with the dataset.☆16Dec 3, 2021Updated 4 years ago
- sklearn, tensorflow, random-forest, adaboost, decision-tress, polynomial-regression, g-boost, knn, extratrees, svr, ridge, bayesian-ridge☆11Jul 13, 2023Updated 3 years ago
- Deep learning aided iterative detection algorithm for massive overloaded MIMO channels☆13Mar 5, 2019Updated 7 years ago
- Pytorch implementation of Noisy Student Training for Automatic Speech Recognition and Automatic Pronunciation Error Detection problem☆99May 30, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Implementation of Neural Nets for Communications Channel Decoding using Log Likelihood Ratios☆16Nov 19, 2020Updated 5 years ago
- A directory of the top business machine learning vendors☆16May 25, 2021Updated 5 years ago
- View and compare Bible Translations in an innovative interlinear format. Run on Windows or Web.☆15Jul 20, 2026Updated last week
- A simple server for receiving GPS and IO data from a Teltonika RUT955☆10Jan 3, 2019Updated 7 years ago
- contains many of my preliminary research ideas☆34Jan 12, 2026Updated 6 months ago
- A live speech recognition using Facebooks wav2vec 2.0 model.☆379Feb 4, 2024Updated 2 years ago
- FunAudioLLM homepage☆17Dec 11, 2024Updated last year
- mul-BERT, the official score on the SemEval 2010 Task 8 dataset is up to 90.72 (Macro-F1).☆16Jan 11, 2021Updated 5 years ago
- ☆13Mar 2, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- real time japanese speech recognition translator using wav2vec2☆39Jul 25, 2022Updated 4 years ago
- Convenience OpenCV library for Android☆14Aug 31, 2018Updated 7 years ago
- [ICASSP 2022] Improving End-to-End Contextual Speech Recognition with Fine-Grained Contextual Knowledge Selection☆25Jul 14, 2026Updated 2 weeks ago
- Learning Evasion Strategy in Pursuit-Evasion by Deep Q-Network, ICPR2018.☆13Dec 22, 2018Updated 7 years ago
- Go proof of space library☆14Jan 13, 2016Updated 10 years ago
- 白泽说人话,通万物之情,晓天下万物状貌。☆26Jun 27, 2018Updated 8 years ago
- Code for paper "NOLD: A Neural-Network Optimized Low-Resolution Decoder for LDPC Codes, Submitted for an IEEE journal;"☆19Apr 14, 2021Updated 5 years ago