Code and data for: Low Resource Grammatical Error Correction Using Wikipedia Edits (WNUT 2018)
☆17Jul 16, 2024Updated 2 years ago
Alternatives and similar repositories for boyd-wnut2018
Users that are interested in boyd-wnut2018 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Automatic extraction of edited sentences from text edition histories.☆82Feb 14, 2022Updated 4 years ago
- ☆18Jan 8, 2021Updated 5 years ago
- MaxMatch (M^2) Scorer - Evaluation program for grammatical error correction systems.☆156Sep 27, 2022Updated 4 years ago
- Data and code used in the 2015 ACL paper, "Ground Truth for Grammatical Error Correction Metrics"☆56Dec 17, 2017Updated 8 years ago
- Repository of "An Empirical Study of Incorporating Pseudo Data into Grammatical Error Correction" (EMNLP-IJCNLP 2019)☆68Dec 23, 2019Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Fast + Non-Autoregressive Grammatical Error Correction using BERT. Code and Pre-trained models for paper "Parallel Iterative Edit Models …☆233Mar 24, 2023Updated 3 years ago
- ERRor ANnotation Toolkit: Automatically extract and classify grammatical errors in parallel original and corrected sentences.☆468May 28, 2026Updated 3 months ago
- Stronger Baselines for Grammatical Error Correction Using a Pretrained Encoder-Decoder Model.☆37Apr 6, 2023Updated 3 years ago
- ☆30May 8, 2020Updated 6 years ago
- GMEG☆33Nov 21, 2024Updated last year
- This repository contains the code for applying One-Token Approximation to a pretrained language model using subword-level tokenization.☆12May 7, 2020Updated 6 years ago
- Official implementation of the papers "GECToR – Grammatical Error Correction: Tag, Not Rewrite" (BEA-20) and "Text Simplification by Tagg…☆976May 21, 2024Updated 2 years ago
- Randomly sample lines from massive text files efficiently☆16Apr 1, 2015Updated 11 years ago
- Load subtitles into Netflix☆12Mar 6, 2021Updated 5 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- GrammarTagger — A Neural Multilingual Grammar Profiler for Language Learning☆33Apr 8, 2021Updated 5 years ago
- cLang-8 is a dataset for grammatical error correction.☆113Jul 19, 2022Updated 4 years ago
- Predict edit intentions on Wikipedia☆19Jan 24, 2019Updated 7 years ago
- 2018 Duolingo Shared Task on Second Language Acquisition Modeling (SLAM) (http://sharedtask.duolingo.com/)☆12May 31, 2018Updated 8 years ago
- A dataset of atomic wikipedia edits containing insertions and deletions of a contiguous chunk of text in a sentence. This dataset contai…☆106May 6, 2019Updated 7 years ago
- Python tools to retrieve text from CommonCrawl WARC files based on cdx index.☆18Feb 18, 2022Updated 4 years ago
- A PyTorch implementation of "Reaching Human-level Performance in Automatic Grammatical Error Correction: An Empirical Study"☆50Dec 17, 2018Updated 7 years ago
- The dataset and statistical analysis code released with the submission of EMNLP 2017 paper "Why We Need New Evaluation Metrics for NLG"☆19Nov 16, 2021Updated 4 years ago
- Python API to TalkBankDB.☆13Jan 22, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- récriture inclusive des textes en ligne☆15Dec 19, 2022Updated 3 years ago
- ☆14Jan 4, 2021Updated 5 years ago
- ☆13Mar 1, 2019Updated 7 years ago
- Rust python bindings for symspell☆21Dec 25, 2023Updated 2 years ago
- Heroes of the Storm data in json format☆15Jul 2, 2026Updated 2 months ago
- ☆34Jul 4, 2018Updated 8 years ago
- Dataset of spoken conversational search utterances☆14Aug 27, 2021Updated 5 years ago
- Larger-Context NMT☆13Aug 20, 2017Updated 9 years ago
- Improved version of GECToR☆62Jul 24, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Neural models and instructions on how to reproduce our results for our neural grammatical error correction systems from M. Junczys-Dowmun…☆88Jun 13, 2019Updated 7 years ago
- The repository of EMNLP 2023 "MixEdit: Revisiting Data Augmentation and Beyond for Grammatical Error Correction"☆12Nov 25, 2023Updated 2 years ago
- Reverse engineer patterns for use with SpaCy's DependencyMatcher☆36Feb 8, 2020Updated 6 years ago
- ☆32Jun 16, 2021Updated 5 years ago
- Code & Data for our Paper "RobustGEC: Robust Grammatical Error Correction Against Subtle Context Perturbation" (EMNLP 2023)☆17Jan 23, 2024Updated 2 years ago
- Scripts for finetuning m2m-100 models☆19Jul 28, 2022Updated 4 years ago
- Examples for using the SiLLM framework for training and running Large Language Models (LLMs) on Apple Silicon☆16May 8, 2025Updated last year