Speech to Text with Hugging Face and Wav2vec 2.0
☆35Feb 13, 2021Updated 5 years ago
Alternatives and similar repositories for speech-to-text
Users that are interested in speech-to-text are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Record GPU memory accesses of a CUDA program and visualize the access pattern in a browser☆13Nov 17, 2020Updated 5 years ago
- Accelerating CNN's convolution operation on GPUs by using memory-efficient data access patterns.☆14Dec 8, 2017Updated 8 years ago
- Speech recognition with federated learning☆11Jan 9, 2020Updated 6 years ago
- a simplified version of wav2vec(1.0, vq, 2.0) in fairseq☆170Sep 21, 2020Updated 5 years ago
- Self-Supervised Speech/Sound Pre-training and Representation Learning Toolkit☆13Nov 18, 2022Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A framework for evaluating the effectiveness of chain-of-thought reasoning in language models.☆19Feb 6, 2025Updated last year
- Minimal implementation of Contrastive Predictive Coding for audio.☆18Nov 17, 2019Updated 6 years ago
- Hierarchical annotation - line (phrase), syllable, phoneme annotations of the jingju (Beijing opera) a-cappella singing dataset☆21Mar 1, 2017Updated 9 years ago
- This repository will contain links to the most famous available books of ML that are online☆13Oct 15, 2024Updated last year
- Protecting Real-Time GPU Kernels on Integrated CPU-GPU SoC Platforms☆12Apr 9, 2018Updated 8 years ago
- This is the public repository for SALSA-Lite features for polyphonic sound event localization and detection using microphone arrays.☆15Dec 3, 2021Updated 4 years ago
- Introducción a la ciencia de datos y al aprendizaje automático☆10Nov 2, 2017Updated 8 years ago
- Deep Complex UNet for speech enhancement, init from "https://github.com/chanil1218/DCUnet.pytorch"☆13Feb 21, 2020Updated 6 years ago
- Local Action, Global Impact (Selected as Top 50 in the 2022 Solution Challenge.)☆17Jan 18, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆45Dec 15, 2022Updated 3 years ago
- Welcome to the Real-Time Voice Activity Detection (VAD) program, powered by Silero-VAD model! 🚀 This program allows you to perform live …☆12Jul 9, 2023Updated 3 years ago
- ☆11Jan 12, 2026Updated 6 months ago
- Starter Astronot for angular website template themes☆11Apr 6, 2024Updated 2 years ago
- open-vocabulary sound event detection☆54Dec 17, 2025Updated 7 months ago
- Pretrained spoken language classifiers from audio.☆10Jan 21, 2021Updated 5 years ago
- Visual Hash for matching copies of visually similar images.☆16Mar 17, 2025Updated last year
- A PyTorch implementation of "Self-Supervised GNN that Jointly Learns to Augment" or "Jointly Learnable Data Augmentations for Self-Superv…☆13Dec 13, 2021Updated 4 years ago
- Universal differential equations for ecologists☆17Apr 24, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆14Sep 20, 2023Updated 2 years ago
- DUSTED: Spoken-Term Discovery using Discrete Speech Units☆17Oct 2, 2024Updated last year
- Various algorithms for voice activity detection☆22Jan 31, 2017Updated 9 years ago
- A Lionel Messi drawing in Python using sketchpy library as a tribute for winning the 2022 FIFA World Cup with Argentina team.☆15Mar 5, 2024Updated 2 years ago
- smart task flow for ops dev workflow☆10Sep 27, 2023Updated 2 years ago
- Collaborative markdown with math☆13Sep 16, 2014Updated 11 years ago
- Introduction Note: This edition of the book is the same as The Rust Programming Language available in print and ebook format from No…☆13Oct 30, 2019Updated 6 years ago
- Tr-VAD: An Efficient Transformer based Voice Activity Detection Model☆18Aug 1, 2024Updated 2 years ago
- Most basic AI Assistant demo derived from the DeepPavlov Dream AI Assistant.☆14May 22, 2023Updated 3 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- We present a deep learning approach towards the large-scale prediction and analysis of bird acoustics from 100 different bird species☆21Jun 17, 2024Updated 2 years ago
- Speech Recognition for speakers with speech disorders due to diseases like Cerebral Palsy, Parkinson or Amyotrophic Lateral Sclerosis ALS…☆23Mar 26, 2017Updated 9 years ago
- Using Extractive summarization to summarize medium posts☆11Nov 17, 2019Updated 6 years ago
- statically generated weekly digest of articles read in Pocket☆10May 14, 2019Updated 7 years ago
- A nuxt module to expose Vuex state in the browser URL for easy sharing☆12Aug 28, 2017Updated 8 years ago
- ☆13May 13, 2017Updated 9 years ago
- Pytorch implementation of MICCAI-2022 paper, Domain-adaptive 3D Medical Image Synthesis: An Efficient Unsupervised Approach https://arxiv…☆22Jul 5, 2022Updated 4 years ago