Modify Chinese text, modified on LaserTagger Model. 文本复述,基于lasertagger做中文文本数据增强。
☆319Jan 3, 2024Updated 2 years ago
Alternatives and similar repositories for text_data_enhancement_with_LaserTagger
Users that are interested in text_data_enhancement_with_LaserTagger are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Modify Chinese text, modified on LaserTagger Model. I name it "文本手术刀".目前,本项目实现了一个文本复述任务,用于NLP语料的数据增强。☆215Mar 24, 2023Updated 3 years ago
- lasertagger-chinese;lasertagger中文学习案例,案例数据,注释,shell运行☆75Mar 25, 2023Updated 3 years ago
- ☆603Mar 12, 2026Updated 5 months ago
- 一键中文数据增强包 ; NLP数据增强、bert数据增强、EDA:pip install nlpcda☆1,876Mar 18, 2025Updated last year
- Research on the Construction and Application of Paraphrase Parallel Corpus☆11Oct 26, 2020Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- An implement of the paper of EDA for Chinese corpus.中文语料的EDA数据增强工具。NLP数据增强。论文阅读笔记。☆1,382May 31, 2022Updated 4 years ago
- 基于bert进行中文文本纠错☆242Jun 12, 2023Updated 3 years ago
- PMI, 是互信息(NMI)中的一种特例, 而互信息,是源于信息论中的一个概念,主要用于衡量2个信号的关联程度.至于PMI,是在文本处 理中,用于计算两个词语之间的关联程度.比起传统的相似度计算, pmi的好处在于,从统计的角度发现词语共现的情况来分析出词语间是否存在语义相关…☆15Aug 24, 2020Updated 6 years ago
- Keyphrase or Keyword Extraction 基于预训练模型的中文关键词抽取方法(论文SIFRank: A New Baseline for Unsupervised Keyphrase Extraction Based on Pre-trained La…☆430May 17, 2020Updated 6 years ago
- 高质量中文预训练模型集合:最先进大模型、最快小模型、相似度专门模型☆811Jul 8, 2020Updated 6 years ago
- 天池 疫情相似句对判定大赛 线上第一名方案☆434Oct 17, 2020Updated 5 years ago
- Open Language Pre-trained Model Zoo☆1,003Nov 18, 2021Updated 4 years ago
- Open Source Pre-training Model Framework in PyTorch & Pre-trained Model Zoo☆3,112May 9, 2024Updated 2 years ago
- A PyTorch-based knowledge distillation toolkit for natural language processing☆1,709May 8, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Pre-trained Chinese ELECTRA(中文ELECTRA预训练模型)☆1,434Apr 19, 2026Updated 4 months ago
- Language Understanding Evaluation benchmark for Chinese: datasets, baselines, pre-trained models,corpus and leaderboard☆1,782Feb 18, 2023Updated 3 years ago
- Pre-Training with Whole Word Masking for Chinese BERT(中文BERT-wwm系列模型)☆10,226Apr 19, 2026Updated 4 months ago
- a bert for retrieval and generation