Python package and cli tool to convert wave files (WAV or AIFF) to vector graphics (SVG, PostScript, CVS)
☆111Jan 3, 2026Updated 8 months ago
Alternatives and similar repositories for wav2vec
Users that are interested in wav2vec are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Enable RNNLM lattice rescoring with Pytorch [kaldi]☆12Jun 5, 2020Updated 6 years ago
- ☆11Oct 3, 2021Updated 4 years ago
- Attention-based model for keywords spotting☆19Aug 9, 2021Updated 5 years ago
- Jasper 기반 양자화된 모델인 Quartznet 한국어 음성인식☆22Jul 21, 2021Updated 5 years ago
- CTC Decoder implementation with python only. Also supports language model decoding using KenLM.☆37May 3, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- An implementation of RNN-Transducer loss in TF-2.0.☆46Jan 7, 2026Updated 8 months ago
- Simple Python library, distributed via binary wheels with few direct dependencies, for easily using wav2vec 2.0 models for speech recogni…☆23Aug 16, 2021Updated 5 years ago
- Repo accompanying the blog post "How to Deploy A State-of-the-art PyTorch Model to iOS via Core ML (Part 3)".☆16Jul 3, 2020Updated 6 years ago
- Filtering and Noise Adding Tool☆29May 27, 2022Updated 4 years ago
- QGIS svg icons - Animal silhouettes☆14Nov 17, 2018Updated 7 years ago
- Installable font package for BC Sans.☆27Sep 4, 2026Updated 3 weeks ago
- Companion repository for the paper "A Comparison of Metric Learning Loss Functions for End-to-End Speaker Verification" published at SLSP…☆61Oct 7, 2020Updated 5 years ago
- ☆19Jan 29, 2023Updated 3 years ago
- Wav2kws is keyword spotting (KWS) based on Wav2Vec 2.0. This model shows state-of-the-art in Google Speech Commands datasets V1 and V2.☆13Jun 11, 2021Updated 5 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Code for the paper: Deep Residual Networks with Auditory Inspired Features for Robust Speech Recognition.☆21Mar 22, 2017Updated 9 years ago
- PyTorch implementation of a Time Delay Neural Network (TDNN)☆41Jun 6, 2019Updated 7 years ago
- 한국어 다중분류 감성분석☆20Jun 7, 2022Updated 4 years ago
- An extension of thu-spmi/CAT which contains a full-fledged implementation of CTC-CRF for Tensorflow.☆12Jul 5, 2021Updated 5 years ago
- A JavaScript voice control library based on Mozilla DeepSpeech☆19Jan 5, 2023Updated 3 years ago
- Speech command recognition with capsule network & various NNs / KWS on Google Speech Command Dataset.☆25Jan 28, 2019Updated 7 years ago
- NISQA - Non-Intrusive Speech Quality and TTS Naturalness Assessment☆16Apr 13, 2022Updated 4 years ago
- A Fast Sequence Transducer Implementation with PyTorch Bindings☆200Sep 20, 2022Updated 4 years ago
- Easily deal with Timecode SMPTE format in Javascript☆13Dec 10, 2022Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Chatbot using reinforcement learning☆19May 2, 2019Updated 7 years ago
- open source knowledge for Syllabics font design and development☆10Nov 13, 2024Updated last year
- [Unofficial] Kakaotrans: Kakao translate API for python☆16Mar 29, 2020Updated 6 years ago
- OpenSCAD Pulley Library to create various pulleys and step pulleys in 3D☆12Dec 28, 2021Updated 4 years ago
- GUI applikation for the Klatt formant synthesizer package☆13Jun 26, 2026Updated 3 months ago
- 🫠 check your data, before you wreck your model☆16Aug 11, 2022Updated 4 years ago
- 将百度DeepSpeech的keras后端由theano改为tensorflow,整合mozilla解码模块进行中文语音识别模型部署☆10Dec 2, 2019Updated 6 years ago
- Mother Tongues Dictionaries dictionary creation tool☆15May 21, 2024Updated 2 years ago
- Extract and find/replace text based on arbitrary correspondences while preserving original file formatting. This library is a fork from t…☆11Sep 8, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Implementation for paper "iMetricGAN: Intelligibility Enhancement for Speech-in-Noise using Generative Adversarial Network-based Metric L…☆56Jul 6, 2023Updated 3 years ago
- a simplified version of wav2vec(1.0, vq, 2.0) in fairseq☆171Sep 21, 2020Updated 6 years ago
- ☆14Jan 16, 2026Updated 8 months ago
- kogpt를 oslo로 파인튜닝하는 예제.☆23Aug 26, 2022Updated 4 years ago
- CTC+Beam_Search+kenlm 是用于以汉字为声学模型建模单元的解码系统☆49Jun 27, 2018Updated 8 years ago
- Dialogue generation models (GPT-2 and Meena) of Pingpong, ScatterLab.☆21Nov 15, 2021Updated 4 years ago
- Megatron LM 11B on Huggingface Transformers☆28Jul 11, 2021Updated 5 years ago