Repository for Vajjala & Lucic (2018)
☆73Feb 15, 2024Updated 2 years ago
Alternatives and similar repositories for OneStopEnglishCorpus
Users that are interested in OneStopEnglishCorpus are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Exploring the idea of a generic, language agnostic, CEFR level classifier☆23Apr 13, 2018Updated 8 years ago
- Code used for the paper "Linguistic Features for Readability Assessment" (Deutsch, Jasbi, and Shieber 2020)☆25Jul 19, 2021Updated 5 years ago
- The repository contains the dataset and the code of the paper: Document-Level Text Simplification: Dataset, Metric and Model.☆25Jun 2, 2023Updated 3 years ago
- ☆26Sep 14, 2025Updated last year
- ☆17Mar 2, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Java Bindings for the C++ library DeepSpeech☆10Jun 4, 2020Updated 6 years ago
- ☆10Oct 3, 2023Updated 2 years ago
- This repository contains the two datasets introduced in the paper "Making Science Simple: Corpora for the Lay Summarisation of Scientific…☆28May 13, 2024Updated 2 years ago
- ☆26May 13, 2024Updated 2 years ago
- ☆26May 4, 2022Updated 4 years ago
- Controllable Sentence Simplification with T5☆18May 24, 2023Updated 3 years ago
- Repository for the CommonLit Ease of Readability Corpus☆26Apr 17, 2024Updated 2 years ago
- Code to reproduce the experiments from the paper.☆104Oct 10, 2023Updated 2 years ago
- ☆14Jun 12, 2015Updated 11 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- A collection of text simplification datasets and other resources☆51Sep 10, 2024Updated 2 years ago
- Code for SLT 2016 paper on Grapheme-to-Phoneme conversion using attention based encoder-decoder models☆15Feb 20, 2019Updated 7 years ago
- TextComplexityDE dataset consists of 1000 sentences in the German language with subjective complexity rating, collected from German learn…☆13Apr 8, 2022Updated 4 years ago
- phonetic similarity algorithms☆13Jun 19, 2018Updated 8 years ago
- a python package for cleaning Gutenberg books and dataset☆37May 2, 2025Updated last year
- An implementation of data augmentation methods for natural language processing tasks.☆13Jul 25, 2024Updated 2 years ago
- Reference-less Quality Estimation of Text Simplification Systems☆52Dec 14, 2023Updated 2 years ago
- Implementation of different noise embeddings for noise aware training of Kaldi acoustic models.☆13Feb 13, 2021Updated 5 years ago
- PassivePy: A Tool to Automatically Identify Passive Voice in Big Text Data☆23Mar 6, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆120Sep 9, 2020Updated 6 years ago
- 24-hour Automatic Speech Recognition☆27Jun 4, 2021Updated 5 years ago
- Using YouTube to prepare a speech recognition dataset for any language☆10Mar 30, 2021Updated 5 years ago
- Optimizing Deeper Transformers on Small Datasets https://arxiv.org/abs/2012.15355☆16Nov 2, 2022Updated 3 years ago
- This repo contains the baseline model recipes and pre-trained model for GramVanni hindi ASR challenge☆16Mar 26, 2022Updated 4 years ago
- Agile reading group that works☆13Feb 2, 2022Updated 4 years ago
- 👄🇧🇷 Alinhamento fonético forçado em Português Brasileiro☆13Jul 18, 2025Updated last year
- MirasVoice is a data set consisting speech samples from bilinguals to train neural network for optimization of speaker verification algor…☆19Mar 15, 2020Updated 6 years ago
- Tool for the Automatic Analysis of Syntactic Sophistication and Complexity☆32Nov 4, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 📖 LanMIT: A Toolkit for Improving Language Models in Low-resourced Speech Recognition based on Kaldi.☆22Jul 12, 2019Updated 7 years ago
- Alignment and annotation for comparable documents.☆22Oct 16, 2018Updated 7 years ago
- Feature extraction for accented-speech or pathological speech☆18Apr 2, 2019Updated 7 years ago
- ☆30Nov 23, 2021Updated 4 years ago
- Python version for Doug Biber's Multidimensional Analysis (MDA)☆42May 24, 2026Updated 4 months ago
- Implementation of Multi speaker TTS☆50Jan 2, 2021Updated 5 years ago
- A simple toolkit for conducting analyses using corpus methods☆28Nov 11, 2021Updated 4 years ago