总结OCR领域的主流公开数据集,包含检测&识别、各种场景、各种语言的数据集,并提供数据集的相关信息及下载链接。
☆43Aug 21, 2022Updated 3 years ago
Alternatives and similar repositories for OCR-Datasets
Users that are interested in OCR-Datasets are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆10Nov 22, 2022Updated 3 years ago
- [AAAI'23 Oral] DPText-DETR: Towards Better Scene Text Detection with Dynamic Points in Transformer☆205Aug 31, 2023Updated 2 years ago
- Official implementation of SPTS: Single-Point Text Spotting (ACM MM 2022 Oral)☆145Jul 26, 2023Updated 3 years ago
- DatasetImgLabeler is a image annotation tool for researchers to prepare datasets in ICDAR2015 format☆12Dec 7, 2019Updated 6 years ago
- Improved Text recognition algorithms on different text domains like scene text, handwritten, document, Chinese/English, even ancient book…☆80Feb 4, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆15Feb 28, 2022Updated 4 years ago
- (CVPR 2024) Bridging the Gap Between End-to-End and Two-Step Text Spotting.☆75Jun 11, 2024Updated 2 years ago
- cross-species analysis of cell identities, markers and regulations☆13Jul 11, 2025Updated last year
- ☆13Jun 10, 2025Updated last year
- ☆12Aug 15, 2024Updated 2 years ago
- A Pytorch implementing of A Deep Learning approach to Template Matching. Usie Hypernet + VGG to match the templates.☆13Dec 18, 2021Updated 4 years ago
- ☆18May 10, 2023Updated 3 years ago
- 【2024 ECAI】First Creating Backgrounds Then Rendering Texts: A New Paradigm for Visual Text Blending☆14Jun 16, 2025Updated last year
- A deep learning framework for cervical cancer detection to allow improved accuracy for PAP smear test results☆19Jan 18, 2020Updated 6 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- DCNN Based Image Edge Detection (FCN + HED + UNet + ResNet)☆13Apr 21, 2018Updated 8 years ago
- ☆15Nov 26, 2023Updated 2 years ago
- This repo contains the code for "Table Structure Extraction with Bi-directional Gated Recurrent Unit Networks", ICDAR 2019..☆19Jul 13, 2023Updated 3 years ago
- ACL'2024-Main: Synergetic Event Understanding: A Collaborative Approach to Cross-Document Event Coreference Resolution with Large Languag…☆12Sep 19, 2025Updated 10 months ago
- tensorflow implementation of GHM-C Loss☆12Apr 26, 2019Updated 7 years ago
- 把VIA标注软件导出的json格式文件转换成COCO标注格式,检验并显示coco格式的bbox和mask☆10Jan 1, 2020Updated 6 years ago
- Pytorch re-implementation of Paper: SwinTextSpotter: Scene Text Spotting via Better Synergy between Text Detection and Text Recognition (…☆290Nov 29, 2024Updated last year
- ODM: A Text-Image Further Alignment Pre-training Approach for Scene Text Detection and Spotting☆46Apr 11, 2025Updated last year
- 切割手寫 png 打包 ttf☆14Aug 4, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ABINet++: Autonomous, Bidirectional and Iterative Language Modeling for Scene Text Spotting☆90Feb 11, 2023Updated 3 years ago
- TextAdaIN: Paying Attention to Shortcut Learning in Text Recognizers☆21Jul 26, 2022Updated 4 years ago
- ☆18Mar 19, 2021Updated 5 years ago
- ADVERSARIAL LEARNING FOR SEMI-SUPERVISED SEMANTIC SEGMENTATION tensorflow☆14Apr 29, 2021Updated 5 years ago
- 利用Swin-Unet(Swin Transformer Unet)实现对文档图片里表格结构的识别,Swin-unet (Swin Transformer Unet) is used to identify the document table structure☆27Feb 23, 2024Updated 2 years ago
- HyperCUT: Video Sequence from a Single Blurry Image using Unsupervised Ordering (CVPR'23)☆14Nov 4, 2025Updated 9 months ago
- Implementation of our paper "Global Localization in Large-scale Point Clouds via Roll-pitch-yaw Invariant Place Recognition and Low-overl…☆10Nov 25, 2023Updated 2 years ago
- Implementation of Practical Facial Landmark Detector (PFLD) on Pytorch☆14Jul 23, 2023Updated 3 years ago
- [NeurIPS2021] BOVText: A Large-Scale, Multidimensional Multilingual Dataset for Video Text Spotting☆70Oct 9, 2023Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- resources for text detection, text recognition, and end to end text spotting☆13Apr 23, 2023Updated 3 years ago
- Official Implementation of LatentSwap:An Efficient Latent Code Mapping Framework for Face Swapping☆29Mar 21, 2025Updated last year
- Implementation of paper: CLIPFont: Texture Guided Vector WordArt Generation☆18Oct 8, 2022Updated 3 years ago
- A program for document recognition: from image to text☆29Aug 13, 2018Updated 8 years ago
- a pytorch implement of Adversarially Adaptive Normalization for Single Domain Generalization☆15Jul 25, 2023Updated 3 years ago
- SimpleVQA: Multimodal Factuality Evaluation for Multimodal Large Language Models☆15Feb 20, 2025Updated last year
- This is an official implementation for "FuseAnyPart: Diffusion-Driven Facial Parts Swapping via Multiple Reference Images"☆21May 19, 2025Updated last year