用于生成文本纠错模型(如Gector)需要的大量数据。
☆15Jan 5, 2023Updated 3 years ago
Alternatives and similar repositories for error_text_gen
Users that are interested in error_text_gen are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 基于seq2edit (Gector) 的中文文本纠错。☆29Nov 15, 2022Updated 3 years ago
- A Trie data structure that allows for fuzzy string matching☆11May 24, 2015Updated 11 years ago
- A faster, simpler and distributed implementation of GECToR, a seq2edit GEC model☆16Oct 10, 2022Updated 3 years ago
- ☆16Sep 4, 2019Updated 6 years ago
- MuCGEC中文纠错数据集及文本纠错SOTA模型开源;Code & Data for our NAACL 2022 Paper "MuCGEC: a Multi-Reference Multi-Source Evaluation Dataset for Chinese Gr…☆570Jun 9, 2023Updated 3 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Code & Data for our Paper "NaSGEC: Multi-Domain Chinese Grammatical Error Correction for Native Speaker Texts" (ACL 2023 Findings)☆96Feb 18, 2025Updated last year
- 分享一些S2S在实际应用中遇到的问题和解决方法。☆28Aug 3, 2020Updated 5 years ago
- Source code for the paper "C-LLM: Learn to Check Chinese Spelling Errors Character by Character"☆30Nov 19, 2024Updated last year
- pytorch版损失函数,改写自科学空间文章,【通过互信息思想来缓解类别不平衡问题】、【将“softmax+交叉熵”推广到多标签分类问题】☆12Aug 22, 2021Updated 4 years ago
- Extending NERDA Library for Continual Learning☆11Mar 31, 2024Updated 2 years ago
- 实验苏神的CoSENT的Torch实现☆33Jan 8, 2022Updated 4 years ago
- This repository open-sources our GEC system submitted by THU KELab (sz) in the CCL2023-CLTC Track 1: Multidimensional Chinese Learner Tex…☆15Nov 25, 2023Updated 2 years ago
- ☆29Mar 18, 2020Updated 6 years ago
- A mesh system for adapting multiple large language models.☆11Mar 20, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- NLP/ML面试各类资料链接 汇总(主要Github收集)☆11Mar 3, 2020Updated 6 years ago
- “达观杯”长文本智能处理挑战赛。达观数据提供了一批长文本数据和分类信息,希望选手动用自己的智慧,结合当下最先进的NLP和人工智能技术,深入分析文本内在结构和语义信息,构建文本分类模型,实现精准分类。☆10Jul 20, 2018Updated 8 years ago
- sentence-transformers to onnx 让sbert模型推理效率更快☆166Mar 11, 2022Updated 4 years ago
- The Corpus & Code for EMNLP 2022 paper "FCGEC: Fine-Grained Corpus for Chinese Grammatical Error Correction" | FCGEC中文语法纠错语料及STG模型☆123Apr 12, 2026Updated 3 months ago
- 非官方的MDCSpell论文的实现☆18Oct 16, 2022Updated 3 years ago
- Set up an async pipeline in python using Celery, RabbitMQ and MongoDB. This repo covers the end to end deployment of an async pipeline fo…☆13Sep 23, 2022Updated 3 years ago
- An implementation of transformer-based language model for sentence rewriting tasks such as summarization, simplification, and grammatical…☆28Jul 25, 2024Updated last year
- 文本数据增强☆15Apr 10, 2020Updated 6 years ago
- 针对NER领域提供从线下训练到线上部署的一整套闭环流程☆14Jun 16, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Hybrid RT DETR: Hybrid encoder-decoder network for end-to-end object detection in UAV imagery☆16May 22, 2024Updated 2 years ago
- 基于依存句法与语义角色标注的三元组抽取☆11Sep 6, 2018Updated 7 years ago
- An example implementatation of synchronized queue for inter-process communication in shared memory☆13Feb 17, 2017Updated 9 years ago
- multi task learning for multi-classification using keras☆13Feb 10, 2020Updated 6 years ago
- 根据维基百科历史编辑数据提取纠错语料。☆12Apr 6, 2022Updated 4 years ago
- Code for KDD CUP 2019 Auto-ML track☆21Jul 25, 2019Updated 6 years ago
- 自然语言处理之中文文本分类(以垃圾短信识别为例)☆24Jun 4, 2020Updated 6 years ago
- cpp inference for EmotiVoice☆16Jan 1, 2024Updated 2 years ago
- TensorRT☆11Sep 22, 2020Updated 5 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code for ACL 2023 paper "Learning 'O' Helps for Learning More: Handling the Unlabeled Entity Problem for Class-incremental NER"☆10Jul 17, 2023Updated 3 years ago
- Chinese character variant converter. 中文异体字转换器。☆24Updated this week
- Yet Another Chinese Spelling Check Dataset (YACSC)☆23Oct 25, 2023Updated 2 years ago
- MedDistant19: Towards an Accurate Benchmark for Broad-Coverage Biomedical Relation Extraction (COLING 2022)☆19Oct 13, 2022Updated 3 years ago
- 抽取式NLP模型(阅读理解模型,MRC)实现词义消歧(WSD)☆14May 10, 2022Updated 4 years ago
- The Code & Paper for ACL 2023 paper "Enhancing Language Representation with Constructional Information for Natural Language Understanding…☆20Jan 18, 2025Updated last year
- Source code for ACM TOIS 2017 paper "A Neural Network Approach to Joint Modeling Social Networks and Mobile Trajectories".☆12Jun 18, 2019Updated 7 years ago