A Multi-Modal Dataset of Chinese Governmental Docunments
☆45Dec 8, 2020Updated 5 years ago
Alternatives and similar repositories for GovDoc-CN
Users that are interested in GovDoc-CN are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 政务公文知识图谱构建☆22Oct 12, 2022Updated 3 years ago
- Testing DeepSpeed integration in 🤗 Accelerate☆12Jun 28, 2022Updated 4 years ago
- [TALLIP] General and Domain Adaptive Chinese Spelling Check with Error Consistent Pretraining☆65Aug 10, 2026Updated 3 weeks ago
- code and data for "CSCD-NS: a Chinese Spelling Check Dataset for Native Speakers"☆86Aug 18, 2024Updated 2 years ago
- ☆10Jun 14, 2023Updated 3 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- FinCUGE Instruction dataset☆16Apr 29, 2023Updated 3 years ago
- A simple summary of fine-grained sentiment analysis☆12Dec 10, 2021Updated 4 years ago
- This repo contains script using Tesseract OCR to digitize pdf ebooks to text format.☆26Feb 24, 2024Updated 2 years ago
- Code for the paper "Learning Variational Word Masks to Improve the Interpretability of Neural Text Classifiers"☆18Dec 15, 2020Updated 5 years ago
- 这是一份集成了RAG和微调以及思维链的LLM应用!最近也结合了知识图谱以及智能体agent~后续还会有很多更新!☆19Oct 12, 2024Updated last year
- 中文soft-masked bert文本纠错复现☆21May 20, 2021Updated 5 years ago
- 复 习论文《A Frustratingly Easy Approach for Joint Entity and Relation Extraction》☆32May 15, 2021Updated 5 years ago
- crawl the public files of different governments through python 3.☆15Aug 29, 2019Updated 7 years ago
- 宠物托管管理系统☆21Mar 23, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆17Apr 27, 2015Updated 11 years ago
- spring boot common 让在spring boot上开发更简单,开箱即用的web组件、分布式锁组件等各种常用组件☆17Aug 30, 2026Updated last week
- ☆15Feb 9, 2023Updated 3 years ago
- Source code for the paper "C-LLM: Learn to Check Chinese Spelling Errors Character by Character"☆30Nov 19, 2024Updated last year
- Hierarchical Models for long document encoding☆22May 29, 2023Updated 3 years ago
- ☆11May 26, 2020Updated 6 years ago
- Code for the paper Neural Pipeline for Zero-Shot Data-to-Text Generation☆16Aug 26, 2024Updated 2 years ago
- Keras library for building (Universal) Transformers, facilitating BERT and GPT models☆11Jan 29, 2019Updated 7 years ago
- 针对从事行政工作公文写作的小伙伴提供了一个快速调整文档格式的word COM加载项。☆34Dec 23, 2019Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- This xblock allows student taking notes from video☆12Sep 21, 2015Updated 10 years ago
- ☆14Apr 19, 2024Updated 2 years ago
- A Specialist-annotated Dataset for Medical-domain Chinese Spelling Correction☆38Jun 6, 2022Updated 4 years ago
- 把 cubox 稍后读软件的「归档」内容转存到其他地方(如Notion),以突破其只能存200条数据的限制☆11Dec 31, 2024Updated last year
- ☆12Mar 18, 2019Updated 7 years ago
- lshash for python3☆10Mar 21, 2018Updated 8 years ago
- DocBankLoader is a dataset loader for DocBank, and can convert DocBank to the Object Detection models' format.☆24Mar 17, 2021Updated 5 years ago
- 关键词式指定站点新闻爬虫☆17Sep 19, 2020Updated 5 years ago
- Convert captured images to text using BaiduOCR, GoogleOCR, WindowsOCR, tesseractOCR, RapidOCR or Capture2Text, and translate the resultin…☆94Oct 9, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- 基于预训练模型的中文关键词抽取方法(论文SIFRank: A New Baseline for Unsupervised Keyphrase Extraction Based on Pre-trained Language Model 的中文版代码)☆12May 17, 2020Updated 6 years ago
- ☆15Jun 16, 2023Updated 3 years ago
- 自然语言处理入门小项目:根据语料生成宋词;双向最大匹配+Bi-gram实现中文分词;简单的基于Flask的Web UI展示☆13Dec 13, 2018Updated 7 years ago
- GCN use for semi-construct document information extraction.☆21Aug 5, 2023Updated 3 years ago
- Keras implementation of graph convolutional networks for sequence labelling☆12Sep 21, 2018Updated 7 years ago
- XBlock to use SCORM content in Open edX. Main development in use_ssla_player branch, requires commercial SSLA player by JCA Solutions.☆12Jun 21, 2023Updated 3 years ago
- MNBVC项目-ShareGPT语料清洗☆16Oct 4, 2023Updated 2 years ago