Scraping Wikipedia for fair use sentences
☆54Jan 25, 2024Updated 2 years ago
Alternatives and similar repositories for cv-sentence-extractor
Users that are interested in cv-sentence-extractor are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Tool to collect and review sentences for Common Voice☆83May 10, 2023Updated 3 years ago
- Script for bundling Common Voice (https://commonvoice.mozilla.org/) clips by language☆11Apr 13, 2023Updated 3 years ago
- Metadata and versioning details for the Common Voice dataset☆175Updated this week
- A living document for all things Common Voice.☆14Jun 24, 2024Updated 2 years ago
- Tooling for producing French dataset for Common Voice☆101Jan 20, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Mozilla Voice Community Playbook☆49May 21, 2024Updated 2 years ago
- Command line tool to create corpora for Common Voice☆78Mar 25, 2026Updated 5 months ago
- Automatic Speech Recognition (ASR) - Kabyle☆19Nov 28, 2020Updated 5 years ago
- หนังสือ "Interpretable Machine Learning" โดย Christoph Molnar ฉบับแปลภาษาไทย / Thai translation of "Interpretable Machine Learning" book…☆15Oct 15, 2021Updated 4 years ago
- Linguistic processing for Common Voice☆59Jan 18, 2024Updated 2 years ago
- Administrative tools for the Tatoeba website☆16May 2, 2021Updated 5 years ago
- Scripts to simplify data prepping for Mozilla DeepSpeech.☆15Aug 6, 2019Updated 7 years ago
- A webpage and API for using Mozilla DeepSpeech☆48Feb 24, 2021Updated 5 years ago
- Tuddar, ismawen d imeḍqan☆11Jan 3, 2020Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Anaouder mouezh e Brezhoneg gant Vosk☆15Nov 24, 2025Updated 9 months ago
- Evaluation of the classification performance (Speech, Music, and Noise) of 1D (WaveNet) and 2D (MobileNet) CNN and RNN (GRU) on the MUSAN…☆15Sep 23, 2020Updated 5 years ago
- Linking topics and learning resources☆10Jun 30, 2016Updated 10 years ago
- Tooling for producing Italian model (public release available) for DeepSpeech and text corpus☆95Mar 15, 2022Updated 4 years ago
- 製作漢字相關語音字典及輸入法,目的只爲保存各地區言語及文化。☆13Jun 4, 2020Updated 6 years ago
- Whisper finetuned on VinBigdata-VLSP2020-100h + KenLM☆38Oct 6, 2023Updated 2 years ago
- Coqui Inference Engine☆41Aug 3, 2021Updated 5 years ago
- ☆41Jul 28, 2025Updated last year
- Natural language processing for the kabyle language☆16Jul 3, 2020Updated 6 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- 🗓 Paper Reading Schedule of NLP Group☆10Sep 6, 2020Updated 6 years ago
- ☆12Jan 8, 2026Updated 8 months ago
- Multilingual dictionary for constructed languages☆15Apr 1, 2026Updated 5 months ago
- Wikidata Live Changes - Group Project - 2020☆11Apr 23, 2024Updated 2 years ago
- 收集客家語言用字、短語、諺語、歌謠和客家語言拼音。☆15Sep 4, 2017Updated 9 years ago
- Google's TPGST reimplementation.☆34Dec 11, 2019Updated 6 years ago
- Tool for creation, manipulation and maintenance of voice corpora☆82May 3, 2024Updated 2 years ago
- Implementation of different noise embeddings for noise aware training of Kaldi acoustic models.☆13Feb 13, 2021Updated 5 years ago
- Esperanto dictionary Mastodon bot☆11Nov 18, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A library of speech gadgets.☆16Oct 15, 2022Updated 3 years ago
- Scripts to work with Chinese language data☆14Jan 21, 2022Updated 4 years ago
- Common Voice is part of Mozilla's initiative to help teach machines how real people speak.☆3,487Updated this week
- Read-only mirror of https://gitlab.gnome.org/GNOME/gnome-session☆21Updated this week
- Python API for reading and querying ARPA formatted language models.☆33Sep 9, 2014Updated 12 years ago
- Source package for iso-flag-png☆24Jul 21, 2025Updated last year
- Using YouTube to prepare a speech recognition dataset for any language☆10Mar 30, 2021Updated 5 years ago