中文论文、证券类、财报类PDF数据
☆41Jun 13, 2024Updated 2 years ago
Alternatives and similar repositories for ChineseDocumentPDF
Users that are interested in ChineseDocumentPDF are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 360LayoutAnaylsis, a series Document Analysis Models and Datasets deleveped by 360 AI Research Institute☆305Sep 10, 2024Updated 2 years ago
- CTE: Contextualized Table Extraction Dataset☆17Feb 23, 2023Updated 3 years ago
- 使用onnxruntime部署MOWA:多合一图像扭曲模型,能处理6种图像扭曲任务,依然是包含C++和Python两个版本的程序☆34Jul 7, 2024Updated 2 years ago
- 使用ONNXRuntime部署DeDoDe:"局部特征匹配:检测,不要描述——描述,不要检测"。依然是C++和Python两个版本的程序☆23Dec 22, 2023Updated 2 years ago
- [IJCAI-2024] The official code of Self-Supervised Pre-training with Symmetric Superimposition Modeling for Scene Text Recognition☆10Aug 10, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆11Feb 23, 2024Updated 2 years ago
- What Is a Good Caption? A Comprehensive Visual Caption Benchmark for Evaluating Both Correctness and Thoroughness☆28May 16, 2025Updated last year
- onnx-java,这里利用java加载onnx模型,并进行推理。☆22May 19, 2022Updated 4 years ago
- ☆172Sep 14, 2026Updated 3 weeks ago
- 【间隙·树·排序算法】 对OCR结果或PDF提取的文本进行版面分析,按人类阅读顺序进行排序。☆173Feb 28, 2024Updated 2 years ago
- ☆21Feb 16, 2025Updated last year
- CDLA: A Chinese document layout analysis (CDLA) dataset☆295Sep 13, 2021Updated 5 years ago
- [ACM'MM 2024 Oral] Official code for "OneChart: Purify the Chart Structural Extraction via One Auxiliary Token"☆268Apr 14, 2025Updated last year
- Analysis of Chinese and English layouts 中英文版面分析☆280Mar 24, 2026Updated 6 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Open-source Agent SDK for Python. Runs the full agent loop in-process — no CLI required.☆46Apr 3, 2026Updated 6 months ago
- 检测和提取各种场景图片中的表格区域,并纠正透视和旋转问题 Detect and extract table regions from images in various scenarios, and correct perspective and rotation i…☆119Dec 10, 2024Updated last year
- Dataset of PNG images from synthetically generated table layouts with annotations in JSONL files☆155Sep 17, 2025Updated last year
- NAF-DPM: A Nonlinear Activation-Free Diffusion Probabilistic Model for Document Enhancement☆55Aug 5, 2024Updated 2 years ago
- OCRVerse: Towards Holistic OCR in End-to-End Vision-Language Models☆29Feb 4, 2026Updated 8 months ago
- ☆11Jan 27, 2020Updated 6 years ago
- This repository summaries publications on Recognition of Handwritten Mathematical Expressions☆15Oct 27, 2017Updated 8 years ago
- DocBank 文档图像增强数据集,此数据集用于文档图像增强,具体任务包括以下内容:Seal detection & Removal 印章检测 & 移除 ;Watermark detection & Removal 水印检测 & 移除;Document deblurrin…☆51Oct 22, 2024Updated last year
- ☆143Feb 13, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- This project could help to reduce the ibdata1 file size.☆10Dec 31, 2017Updated 8 years ago
- ☆10Mar 6, 2016Updated 10 years ago
- YOLO models trained by DocLayNet - power your Document Intelligent by Layout Analysis☆165Mar 10, 2026Updated 7 months ago
- XFUND: A Multilingual Form Understanding Benchmark☆223Jul 15, 2022Updated 4 years ago
- Repository for ACL2020 paper "Refer360° A Referring Expression Recognition Dataset in 360°Images"☆15Jun 26, 2021Updated 5 years ago
- YOLOv10 trained on DocLayNet dataset.☆81Nov 1, 2024Updated last year
- [TPAMI] Locating and Counting Heads in Crowds With a Depth Prior☆10Jan 7, 2022Updated 4 years ago
- ☆17Jan 5, 2024Updated 2 years ago
- 🎸store guitar tabs☆31Sep 4, 2026Updated last month
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- 整理目前开源的最优表格识别模型,完善前后处理,模型转换为ONNX | Organize the currently open-source optimal table recognition models, improve pre-processing and post-…☆966Aug 3, 2025Updated last year
- A simple algorithm to find ordered key-value pairs from paddleOCR recognition outputs☆10Mar 1, 2021Updated 5 years ago
- The official code of Linguistic More: Taking a Further Step toward Efficient and Accurate Scene Text Recognition (IJCAI2023)☆26Sep 3, 2023Updated 3 years ago
- 🤡 An up-to-date & curated list of awesome KBQA papers, methods & resources.☆10Jul 14, 2022Updated 4 years ago
- Source-Free Domain Adaptation with Contrastive Domain Alignment and Self-supervised Exploration for Face Anti-Spoofing, ECCV2022☆23Nov 28, 2022Updated 3 years ago
- Data and code for ACL 2023 paper "RobuT: A Systematic Study of Table QA Robustness Against Human-Annotated Adversarial Perturbations"☆15Feb 8, 2024Updated 2 years ago
- official code for "Fox: Focus Anywhere for Fine-grained Multi-page Document Understanding"☆198May 31, 2024Updated 2 years ago