Scraping Wikipedia for fair use sentences
☆54Jan 25, 2024Updated 2 years ago
Alternatives and similar repositories for cv-sentence-extractor
Users that are interested in cv-sentence-extractor are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Tool to collect and review sentences for Common Voice☆83May 10, 2023Updated 3 years ago
- Script for bundling Common Voice (https://commonvoice.mozilla.org/) clips by language☆11Apr 13, 2023Updated 3 years ago
- Running Mozilla's implementation of Baidu DeepSpeech on Google Colaboratory☆16Mar 18, 2019Updated 7 years ago
- Metadata and versioning details for the Common Voice dataset☆174Jun 16, 2026Updated 2 months ago
- A JSON dataset of information about language museums around the world☆13Feb 26, 2020Updated 6 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- A living document for all things Common Voice.☆14Jun 24, 2024Updated 2 years ago
- Tooling for producing French dataset for Common Voice☆101Jan 20, 2025Updated last year
- Mozilla Voice Community Playbook☆49May 21, 2024Updated 2 years ago
- Command line tool to create corpora for Common Voice☆78Mar 25, 2026Updated 5 months ago
- หนังสือ "Interpretable Machine Learning" โดย Christoph Molnar ฉบับแปลภาษาไทย / Thai translation of "Interpretable Machine Learning" book…☆15Oct 15, 2021Updated 4 years ago
- Linguistic processing for Common Voice☆59Jan 18, 2024Updated 2 years ago
- Spoken Language Identification on Common Voice and AudioSet using Deep Learning☆42Feb 4, 2026Updated 6 months ago
- Scripts to simplify data prepping for Mozilla DeepSpeech.☆15Aug 6, 2019Updated 7 years ago
- Simple, fast dictionary-based language detector for short texts.☆21Feb 5, 2026Updated 6 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Draftify is a simple note app to write passing thought and ideas without distractions.☆14Mar 9, 2023Updated 3 years ago
- Tuddar, ismawen d imeḍqan☆11Jan 3, 2020Updated 6 years ago
- Whisper finetuned on VinBigdata-VLSP2020-100h + KenLM☆38Oct 6, 2023Updated 2 years ago
- Coqui Inference Engine☆41Aug 3, 2021Updated 5 years ago
- 🌻 MediaWiki extension allowing mass recording of clean, well cut, well named pronunciation files.☆17Updated this week
- Python package to simulate a vast range of transmission processes on various structures☆13Nov 8, 2022Updated 3 years ago
- 🗓 Paper Reading Schedule of NLP Group☆10Sep 6, 2020Updated 5 years ago
- ☆12Jan 8, 2026Updated 7 months ago
- A Unsplash Search Application built with Svelte 3☆17Dec 4, 2022Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Multilingual dictionary for constructed languages☆14Apr 1, 2026Updated 4 months ago
- Wikidata Live Changes - Group Project - 2020☆11Apr 23, 2024Updated 2 years ago
- Ĉi tiu deponejo enhavas la fontokodon de la retejo Telegramo.org. / This repository contains the source code of the website Telegramo.org…☆32Nov 4, 2019Updated 6 years ago
- Implementation of different noise embeddings for noise aware training of Kaldi acoustic models.☆13Feb 13, 2021Updated 5 years ago
- A library of speech gadgets.☆15Oct 15, 2022Updated 3 years ago
- Firefox Voice is an experiment in a voice-controlled web user agent☆292Jan 29, 2021Updated 5 years ago
- Common Voice is part of Mozilla's initiative to help teach machines how real people speak.☆3,486Updated this week
- Python API for reading and querying ARPA formatted language models.☆33Sep 9, 2014Updated 11 years ago
- Source package for iso-flag-png☆23Jul 21, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- 🐸STT integration examples☆132Sep 23, 2022Updated 3 years ago
- Using YouTube to prepare a speech recognition dataset for any language☆10Mar 30, 2021Updated 5 years ago
- A Chainer implementation of a Convolutional Network model for relation classification in the SemEval Task 8 dataset. This model performs …☆17Jan 16, 2018Updated 8 years ago
- ☆33Jan 23, 2024Updated 2 years ago
- Convolutional REpresenations for Music Analysis☆13Jul 5, 2016Updated 10 years ago
- PyTorch speech2text inference script for the NVidia openseq2seq wav2letter model variant☆10Aug 12, 2019Updated 7 years ago
- C++ implementation of End to End TTS which combines both Tacatron2 and LPCNET Vocoder.☆32Oct 1, 2019Updated 6 years ago