Speech to Text with self-supervised learning based on wav2vec 2.0 framework using Hugging Face's Transformer
☆29Jun 1, 2021Updated 5 years ago
Alternatives and similar repositories for wav2vec2-huggingface-demo
Users that are interested in wav2vec2-huggingface-demo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This utility allows one to cut multiple clips from a single or multiple audio files.☆19Apr 13, 2026Updated 4 months ago
- Arabic Phonetic Dictionary Generator Tool for Automatic Speech Recognition Applications☆11Oct 27, 2021Updated 4 years ago
- Project for HIDING SPEAKER’S SEX IN SPEECH USING ZERO-EVIDENCE SPEAKER REPRESENTATION IN AN ANALYSIS/SYNTHESIS PIPELINE☆15Nov 30, 2022Updated 3 years ago
- ☆39Sep 26, 2020Updated 5 years ago
- Introduction to AI bootcamp at KAU☆10Jan 8, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆10Mar 31, 2025Updated last year
- A tool designed to extract numerical data from scanned historical weather documents.☆13Dec 1, 2024Updated last year
- A blockchain-based local energy market (LEM) simulation.☆10Jun 21, 2022Updated 4 years ago
- bm25 is a scoring function that helps with information retrieval☆14Sep 17, 2020Updated 5 years ago
- Goodness of Pronunciation algorithm using PyKaldi☆19Jun 12, 2022Updated 4 years ago
- This repository contains code implementation of the paper "AI-Guardian: Defeating Adversarial Attacks using Backdoors, at IEEE Security a…☆14Aug 13, 2023Updated 3 years ago
- 词、句拼音转汉字、拼音分割、拼音补全、pygame输入中文☆15Mar 21, 2020Updated 6 years ago
- provide SPHERE-formatted output as well as RIFF, AU, AIFF and raw☆14Dec 18, 2021Updated 4 years ago
- ☆11Apr 7, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Towards Learning Multi-domain Crowd Counting (T-CSVT, 2021), https://ieeexplore.ieee.org/document/9658506☆13Apr 6, 2023Updated 3 years ago
- Implement and train a neural network from scratch in Python for the MNIST dataset (no PyTorch).☆14Mar 22, 2021Updated 5 years ago
- Using-Deep-Learning-Techniques-perform-Fracture-Detection-Image-Processing Using Different Image Processing techniques Implementing Fract…☆14May 20, 2022Updated 4 years ago
- ☆13Mar 27, 2020Updated 6 years ago
- Semi-supervised spoken language understanding (SLU) via self-supervised speech and language model pretraining☆12Mar 23, 2021Updated 5 years ago
- ☆11Nov 30, 2021Updated 4 years ago
- simple NMT With Attention For Arabic to English☆11Mar 5, 2022Updated 4 years ago
- ☆12Oct 13, 2022Updated 3 years ago
- An implementation of the paper titled "Arabic Speech Emotion Recognition Employing Wav2vec2.0 and HuBERT Based on BAVED Dataset" https://…☆16Feb 17, 2022Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A live speech recognition using Facebooks wav2vec 2.0 model.☆380Feb 4, 2024Updated 2 years ago
- CropML Python library☆21Updated this week
- docker for HF wav2vec2-sprint☆13Mar 26, 2021Updated 5 years ago
- Code accompanying AES Semantic Audio Conference paper titled "A Dataset and Method for Guitar Solo Detection in Rock Music"☆11Jan 18, 2018Updated 8 years ago
- DJI Phantom 4 Multispectral raw image radiometric calibration☆13Jun 4, 2021Updated 5 years ago
- Crawl & visualize ICLR papers and reviews.☆18Nov 5, 2022Updated 3 years ago
- ☆10Mar 15, 2022Updated 4 years ago
- ☆21Feb 12, 2025Updated last year
- Phase-aware Adversarial Defense for Improving Adversarial Robustness☆11Oct 12, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆12Jul 5, 2023Updated 3 years ago
- J-Net is aimed for audio separation with randomly weighted encoder.☆12Oct 23, 2019Updated 6 years ago
- Presents an optimised Sentinel-1-based Soil Moisture estimation workflow on tilted topography with permanent vegetation cover☆16Aug 18, 2022Updated 4 years ago
- Urban Sound Classification : striving towards a fair comparison☆17Dec 11, 2020Updated 5 years ago
- Decentralized marketplace application built on the Ethereum Blockchain using the Truffle Framework☆15Dec 15, 2023Updated 2 years ago
- ☆25Mar 21, 2025Updated last year
- Spoofing Speaker Verification Systems with Multi-speaker Text-to-speech Synthesis☆11Jun 21, 2022Updated 4 years ago