tianchiguaixia / layoutlmv3-chineseView external linksLinks
该项目是为了使用layoutlmv3针对中文图片训练和推理。 其中主要解决三个问题: 1.数据标准化成可以的训练数据集格式 2.layoutlmv3-base-chinese 分词修改 2.超过512长度的文本切分和滑窗操作
☆63Sep 6, 2024Updated last year
Alternatives and similar repositories for layoutlmv3-chinese
Users that are interested in layoutlmv3-chinese are comparing it to the libraries listed below
Sorting:
- chinese document classification of layoutlmv3 and layoutxlm☆46Oct 25, 2022Updated 3 years ago
- Qwen-WisdomVast is a large model trained on 1 million high-quality Chinese multi-turn SFT data, 200,000 English multi-turn SFT data, and …☆18Apr 12, 2024Updated last year
- Code and data for the paper: DTSM: Toward Dense Table Structure Recognition with Text Query Encoder and Adjacent Feature Aggregator☆12Apr 28, 2024Updated last year
- [MM'2024] PEneo, an effective algorithm for key-value pair extraction from form-like documents, designed for real-world applications.☆41Apr 7, 2025Updated 10 months ago
- ☆68Sep 24, 2023Updated 2 years ago
- 使用Qwen1.5-0.5B-Chat模型进行通用信息抽取任务的微调,旨在: 验证生成式方法相较于抽取式NER的效果; 为新手提供简易的模型微调流程,尽量减少代码量; 大模型训练的数据格式处理。☆15Sep 6, 2024Updated last year
- 通用版面分析 | 中文文档解析 |Document Layout Analysis | layout paser☆48Jun 13, 2024Updated last year
- ☆19Mar 10, 2023Updated 2 years ago
- Table Structure Recognition☆28Jul 25, 2024Updated last year
- 表格结构识别LGPMA推理☆25Nov 17, 2022Updated 3 years ago
- ICDAR 2024 Table OCR Model☆39Updated this week
- 360LayoutAnaylsis, a series Document Analysis Models and Datasets deleveped by 360 AI Research Institute☆306Sep 10, 2024Updated last year
- ☆10Jun 22, 2020Updated 5 years ago
- ☆40Jun 15, 2024Updated last year
- [ICDAR 2023] SelfDocSeg: A self-supervised vision-based approach towards Document Segmentation (Oral)☆42Oct 6, 2023Updated 2 years ago
- CDLA: A Chinese document layout analysis (CDLA) dataset☆288Sep 13, 2021Updated 4 years ago
- TRACE: Table Reconstruction Aligned to Corner and Edges (ICDAR 2023)☆30Mar 13, 2024Updated last year
- [CVPR 2025] DocLayLLM: An Efficient Multi-modal Extension of Large Language Models for Text-rich Document Understanding☆26Dec 18, 2025Updated 2 months ago
- llms related stuff , including code, docs☆13Feb 25, 2025Updated 11 months ago
- DocTr++ in PaddlePaddle☆58Jul 24, 2024Updated last year
- 基于TrOCR + UniMER-1M数据集,训练一个小而美的公式识别模型☆29Jun 23, 2025Updated 7 months ago
- 微调阿里开源的文字检测模型,利用合合识别返回的OCR结果作为初始训练数据,对模型进行优化训练,使其更加适应1万张图片的具体场景,提高文字识别的精度。☆10Dec 9, 2024Updated last year
- ☆102Dec 23, 2024Updated last year
- SPRINT: Script-agnostic Structure Recognition in Tables☆16Mar 26, 2025Updated 10 months ago
- 智谱 Realtime API 接口前端使用样例☆26Sep 3, 2025Updated 5 months ago
- High-Performance Transformers for Table Structure Recognition Need Early Convolutions☆44Apr 3, 2024Updated last year
- ☆42Sep 2, 2023Updated 2 years ago
- This is the official repository of the EMNLP 2023 paper Reading Order Matters: Information Extraction from Visually-rich Documents by Tok…☆18Mar 15, 2024Updated last year
- ☆38Oct 20, 2023Updated 2 years ago
- ☆20Jun 21, 2024Updated last year
- ☆11Updated this week
- 阅读顺序、Layoutreader☆19May 8, 2025Updated 9 months ago
- LLM-MapBook: AI-Powered Maps for Storytelling. Extracts geo-coordinates from books, visualizes on interactive maps, offering immersive st…☆20Dec 29, 2024Updated last year
- Code for AAAI 2023 Paper : “Alignment-Enriched Tuning for Patch-Level Pre-trained Document Image Models”☆18Dec 6, 2022Updated 3 years ago
- 百度网盘AI大赛——图像处理挑战赛:文档图像摩尔纹消除第2名方案☆43Nov 28, 2023Updated 2 years ago
- ☆18Feb 5, 2026Updated last week
- Create Slides with a simple MCP server using Python PPTX library☆32May 20, 2025Updated 8 months ago
- Datasets and Evaluation Scripts for CompHRDoc☆56Feb 25, 2025Updated 11 months ago
- A High-efficiency Open-source Toolkit for Table-to-Latex Task☆275Dec 6, 2025Updated 2 months ago