Speech to Text with self-supervised learning based on wav2vec 2.0 framework using Hugging Face's Transformer
☆29Jun 1, 2021Updated 5 years ago
Alternatives and similar repositories for wav2vec2-huggingface-demo
Users that are interested in wav2vec2-huggingface-demo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Arabic Phonetic Dictionary Generator Tool for Automatic Speech Recognition Applications☆11Oct 27, 2021Updated 4 years ago
- Project for HIDING SPEAKER’S SEX IN SPEECH USING ZERO-EVIDENCE SPEAKER REPRESENTATION IN AN ANALYSIS/SYNTHESIS PIPELINE☆15Nov 30, 2022Updated 3 years ago
- Mitigating Open-Vocabulary Caption Hallucinations (EMNLP 2024)☆19Oct 18, 2024Updated last year
- 针对口语进行时间抽取并标准化☆13Mar 2, 2020Updated 6 years ago
- Introduction to AI bootcamp at KAU☆10Jan 8, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆10Mar 31, 2025Updated last year
- ☆24Feb 16, 2024Updated 2 years ago
- A tool designed to extract numerical data from scanned historical weather documents.☆13Dec 1, 2024Updated last year
- A blockchain-based local energy market (LEM) simulation.☆10Jun 21, 2022Updated 4 years ago
- Goodness of Pronunciation algorithm using PyKaldi☆19Jun 12, 2022Updated 4 years ago
- Mainly on text documents. Implemented a Mini Search Engine using different algorithms and then summaried documents using lexrank.☆11Jan 19, 2018Updated 8 years ago
- Code for paper: "RemovalNet: DNN model fingerprinting removal attack", IEEE TDSC 2023.☆10Nov 27, 2023Updated 2 years ago
- Implement and train a neural network from scratch in Python for the MNIST dataset (no PyTorch).☆14Mar 22, 2021Updated 5 years ago
- AdvSV stands as the first dataset developed specifically for evaluating Speaker Verification (SV) systems against adversarial attacks. I…☆11Nov 21, 2023Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Arabic Real-time Sign Language Translator☆15Oct 29, 2021Updated 4 years ago
- ☆11Nov 30, 2021Updated 4 years ago
- 【Demo】对新闻标题使用TF-IDF向量化和cosine相似度计算完成相似标题推荐☆14Mar 2, 2020Updated 6 years ago
- Some python codes to generate various aggregate particles (e.g. ballistic particle cluster, and cluster-cluster).☆16Aug 22, 2019Updated 7 years ago
- Pure Python MGRS coordinate converter.☆15Nov 23, 2025Updated 10 months ago
- ☆12Oct 13, 2022Updated 3 years ago
- An implementation of the paper titled "Arabic Speech Emotion Recognition Employing Wav2vec2.0 and HuBERT Based on BAVED Dataset" https://…☆17Feb 17, 2022Updated 4 years ago
- CropML Python library☆21Sep 17, 2026Updated last week
- SOMOSPIE (Soil Moisture Spatial Inference Engine) consists of a Jupyter Notebook and a suite of machine learning methods to process input…☆16Aug 26, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- DJI Phantom 4 Multispectral raw image radiometric calibration☆13Jun 4, 2021Updated 5 years ago
- Multimodal Transformer for Korean Sentiment Analysis with Audio and Text Features☆28Sep 7, 2021Updated 5 years ago
- Crawl & visualize ICLR papers and reviews.☆18Nov 5, 2022Updated 3 years ago
- ☆14Apr 6, 2025Updated last year
- ☆10Mar 15, 2022Updated 4 years ago
- Worked with a number of real-world datasets available at http://snap.stanford.edu/ to identify the importance of certain nodes in terms o…☆11Dec 20, 2016Updated 9 years ago
- ☆12Jul 5, 2023Updated 3 years ago
- J-Net is aimed for audio separation with randomly weighted encoder.☆12Updated this week
- Decentralized marketplace application built on the Ethereum Blockchain using the Truffle Framework☆15Dec 15, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- LibriVoc is a new open-source, large-scale dataset for vocoder artifact detection. LibriVoc is derived from the LibriTTS speech corpus, w…☆16Nov 6, 2025Updated 10 months ago
- Classification of audio signals using PyTorch☆13May 19, 2020Updated 6 years ago
- Set of algorithms and structures related to geodesy☆17Jun 4, 2019Updated 7 years ago
- PyTorch Implementation of "Learning Natural Language Inference with LSTM", 2016, S. Wang et al. (https://arxiv.org/pdf/1512.08849.pdf)☆19Dec 23, 2022Updated 3 years ago
- Utilities for working with videos☆13Jul 5, 2025Updated last year
- 利用Python爬取网站近年的政府工作报告,并进行简单的词频分析+词云☆22Aug 13, 2026Updated last month
- Identify the type of news based on headlines and short descriptions☆17Mar 25, 2019Updated 7 years ago