☆75Jul 6, 2020Updated 6 years ago
Alternatives and similar repositories for MTHv2_Datasets_Release
Users that are interested in MTHv2_Datasets_Release are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The Tripitaka Koreana in Han (TKH) Dataset and the Multiple Tripitaka in Han (MTH) Dataset for the research of Chinese character detectio…☆73Sep 23, 2020Updated 5 years ago
- ☆32Updated this week
- ☆36Aug 1, 2026Updated last month
- ☆25Jul 17, 2026Updated 2 months ago
- 粤港澳大湾区(黄埔)国际算法算例大赛-古籍文档图像识别与分析算法比赛 Alphx队源码☆46Mar 16, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆105Aug 1, 2026Updated last month
- HRCenterNet: An Anchorless Approach to Chinese Character Segmentation in Historical Documents☆39Mar 8, 2022Updated 4 years ago
- The SCUT-EPT Dataset for the research of offline handwritten Chinese text recognition (HCTR) in educational documents has been released.☆133Aug 1, 2026Updated last month
- Pytorch implementation for "Implicit Feature Alignment: Learn to Convert Text Recognizer to Text Spotter".☆67Jun 15, 2021Updated 5 years ago
- 網絡上臺灣各種正字表均只能找到常用國字和次常用國字的文字版,即甲表和乙表,無罕用國字丙表,瑞據「異體字字典」附錄之正字表進行整理成文字版,以便參考,「異體字字典」正式六版附錄中說共計29921字,實 則整理出來共計29923字,未知哪個統計數據出錯,期網友能找出錯誤加以改正。☆18Jun 29, 2024Updated 2 years ago
- Synthetic Dataset used in the ICDAR2019 Competition on HArvesting Raw Tables from Infographics (CHART-Infographics)☆24Jun 10, 2026Updated 3 months ago
- Sliding Convolutional Attention Network for Scene Text Recognition☆11Aug 31, 2018Updated 8 years ago
- Official implementation of PageNet (IJCV 2022)☆84Oct 31, 2022Updated 3 years ago
- [ACM MM 2022] Marior: Margin Removal and Iterative Content Rectification for Document Dewarping in the Wild☆27Aug 12, 2022Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆83Oct 7, 2023Updated 2 years ago
- Reproducing the Past: A Dataset for Benchmarking Inscription Restoration (ACM MM'24)☆14Oct 15, 2025Updated 11 months ago
- ☆44Jul 9, 2024Updated 2 years ago
- This is the official Pytorch implementation of Generate Like Experts: Multi-Stage Font Generation by Incorporating Font Transfer Process …☆49Feb 5, 2025Updated last year
- Evaluation of Natural Language Processing (NLP) tools for the Ancient Chinese language☆48Mar 15, 2026Updated 6 months ago
- Pytorch implementation for "Decoupled attention network for text recognition".☆315Jul 25, 2024Updated 2 years ago
- Handwritten mathematical expression dataset☆34Jun 23, 2022Updated 4 years ago
- A PyTorch implementation of "From Two to One: A New Scene Text Recognizer with Visual Language Modeling Network" (ICCV2021)☆104Dec 9, 2021Updated 4 years ago
- Keras Layer implementation of Attention☆10Jun 22, 2020Updated 6 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆69Oct 23, 2020Updated 5 years ago
- 古文语言理解测评基准 Classical Chinese Language Understanding Evaluation Benchmark: datasets, baselines, pre-trained models, corpus and leaderboard☆58Aug 23, 2023Updated 3 years ago
- NLP Training/Teaching Materials with Articut☆17Oct 14, 2024Updated last year
- [NeurIPS 2024] IF-Font: Ideographic Description Sequence-Following Font Generation☆42Mar 15, 2025Updated last year
- classic Chinese punctuate experiment with keras using daizhige(殆知阁古代文献藏书) dataset☆35Dec 8, 2022Updated 3 years ago
- Page Segmentation Code. I'm working with OCRopus and the UW-III data set to test how the page segmentation algorithms work with smaller s…☆20Feb 23, 2013Updated 13 years ago
- A Large Dataset of Historical Japanese Documents with Complex Layouts☆37Jun 23, 2026Updated 2 months ago
- Aggregation Cross-Entropy for Sequence Recognition. CVPR 2019.☆301Dec 9, 2021Updated 4 years ago
- Text detection by pixel link and DSSD☆12Dec 16, 2018Updated 7 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Official repository accompaying the ICDAR 2023 paper☆14Oct 3, 2023Updated 2 years ago
- CNN models for CASIA-HWDB dataset recognition, Implementation of paper <Building Fast and Compact Convolutional Neural Networks for Offli…☆50Jun 7, 2018Updated 8 years ago
- 一個用於在網頁上查詢漢字中古音韻地位的瀏覽器擴展程序。☆15Oct 24, 2025Updated 10 months ago
- [NeurIPS 2022 Spotlight] MSDS: A Large-Scale Chinese Signature and Token Digit String Dataset for Handwriting Verification☆95Jul 17, 2026Updated 2 months ago
- Radical Analysis Network for Learning Hierarchies of Chinese Characters☆58Jun 17, 2020Updated 6 years ago
- 西方学者普遍从汉字部件出发理解汉字,该库给出了中文部件分解的详细说明和数据库。☆15Jul 20, 2023Updated 3 years ago
- AES - Ancient Egyptian Sentences; Corpus of Ancient Egyptian sentences for corpus-linguistic research☆11May 18, 2021Updated 5 years ago