总结OCR领域的主流公开数据集,包含检测&识别、各种场景、各种语言的数据集,并提供数据集的相关信息及下载链接。
☆45Aug 21, 2022Updated 4 years ago
Alternatives and similar repositories for OCR-Datasets
Users that are interested in OCR-Datasets are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The official implementation of Image Quality Assessment for Machines: Paradigm, Large-scale Database, and Models.☆22Nov 3, 2025Updated 10 months ago
- The official project of paper "Visual Text Processing: A Comprehensive Review and Unified Evaluation""☆104Oct 20, 2025Updated 11 months ago
- CNN-based Russian OCR (uni project)☆11Jun 18, 2020Updated 6 years ago
- Official implementation of SPTS: Single-Point Text Spotting (ACM MM 2022 Oral)☆145Jul 26, 2023Updated 3 years ago
- DatasetImgLabeler is a image annotation tool for researchers to prepare datasets in ICDAR2015 format☆12Dec 7, 2019Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Improved Text recognition algorithms on different text domains like scene text, handwritten, document, Chinese/English, even ancient book…☆80Feb 4, 2023Updated 3 years ago
- ☆15Feb 28, 2022Updated 4 years ago
- (CVPR 2024) Bridging the Gap Between End-to-End and Two-Step Text Spotting.☆75Jun 11, 2024Updated 2 years ago
- ☆13Jun 10, 2025Updated last year
- Python library for determination of feed forward controls for nonlinear systems.☆17Apr 3, 2016Updated 10 years ago
- ☆18May 10, 2023Updated 3 years ago
- 【2024 ECAI】First Creating Backgrounds Then Rendering Texts: A New Paradigm for Visual Text Blending☆14Jun 16, 2025Updated last year
- Code and data for the paper: DTSM: Toward Dense Table Structure Recognition with Text Query Encoder and Adjacent Feature Aggregator☆14Apr 28, 2024Updated 2 years ago
- Retinaface pytorch face-pose-detect face-key-point-detect☆37Mar 3, 2021Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ICLR2025] Are Large Vision Language Models Good Game Players?☆13Mar 3, 2025Updated last year
- ☆15Nov 26, 2023Updated 2 years ago
- This repo contains the code for "Table Structure Extraction with Bi-directional Gated Recurrent Unit Networks", ICDAR 2019..☆19Jul 13, 2023Updated 3 years ago
- ACL'2024-Main: Synergetic Event Understanding: A Collaborative Approach to Cross-Document Event Coreference Resolution with Large Languag…☆12Sep 19, 2025Updated last year
- tensorflow implementation of GHM-C Loss☆12Apr 26, 2019Updated 7 years ago
- 把VIA标注软件导出的json格式文件转换成COCO标注格式,检验并显示coco格式的bbox和mask☆10Jan 1, 2020Updated 6 years ago
- ODM: A Text-Image Further Alignment Pre-training Approach for Scene Text Detection and Spotting☆46Apr 11, 2025Updated last year
- ☆18Mar 19, 2021Updated 5 years ago
- Use Fully Convolutional Networks (FCNs) for classifying image pixels as (road pixel, not-road-pixel)☆16Oct 11, 2017Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆16Sep 12, 2023Updated 3 years ago
- real-time circle detection by edge drawing.algorithm provided by Cuneyt Akinlar and Cihan Tonal☆18Feb 24, 2017Updated 9 years ago
- Low-Computation Egocentric Barcode Detector for the Blind☆10Jun 9, 2017Updated 9 years ago
- 利用Swin-Unet(Swin Transformer Unet)实现对文档图片里表格结构的识别,Swin-unet (Swin Transformer Unet) is used to identify the document table structure☆27Feb 23, 2024Updated 2 years ago
- HyperCUT: Video Sequence from a Single Blurry Image using Unsupervised Ordering (CVPR'23)☆14Nov 4, 2025Updated 10 months ago
- Cloned repository from Hugging Face Spaces (CVPR 2022 Demo)☆54Sep 29, 2022Updated 3 years ago
- resources for text detection, text recognition, and end to end text spotting☆12Apr 23, 2023Updated 3 years ago
- [NeurIPS2021] BOVText: A Large-Scale, Multidimensional Multilingual Dataset for Video Text Spotting☆70Oct 9, 2023Updated 2 years ago
- Tflite VX Delegate i.MX Machine Learning☆13Updated this week
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- A program for document recognition: from image to text☆29Aug 13, 2018Updated 8 years ago
- a pytorch implement of Adversarially Adaptive Normalization for Single Domain Generalization☆15Jul 25, 2023Updated 3 years ago
- SimpleVQA: Multimodal Factuality Evaluation for Multimodal Large Language Models☆15Feb 20, 2025Updated last year
- Code to reproduce 'MOCCA: Multi-Layer One-Class Classification for Anomaly Detection'☆10Dec 12, 2021Updated 4 years ago
- In OLHWDB ,you can find the ptts files, this code can help you get the information of the ptts☆11Mar 8, 2022Updated 4 years ago
- Official Implementation of OneNet☆22Oct 16, 2025Updated 11 months ago
- unofficial NTU academic poster latex template.☆16Jul 12, 2021Updated 5 years ago