A PyTorch implementation of DTrOCR: Decoder-only Transformer for Optical Character Recognition
☆205Jul 21, 2026Updated 3 weeks ago
Alternatives and similar repositories for DTrOCR
Users that are interested in DTrOCR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An implementation of "CLIP4STR: A Simple Baseline for Scene Text Recognition with Pre-trained Vision-Language Model".☆153Nov 14, 2025Updated 9 months ago
- ☆62Dec 4, 2023Updated 2 years ago
- ☆44Jul 9, 2024Updated 2 years ago
- [WACV2025] source code of StrDA: https://arxiv.org/abs/2410.09913☆19Apr 15, 2025Updated last year
- ☆92Feb 9, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- (Pattern Recognition) Pytorch implementation of “HTR-VT: Handwritten Text Recognition with Vision Transformer”☆134Jan 22, 2026Updated 6 months ago
- Cross-lingual learning in scene text recognition (ICASSP2024)☆19Sep 29, 2024Updated last year
- [ICCV 2023] Code base for Revisiting Scene Text Recognition: A Data Perspective☆206Nov 1, 2023Updated 2 years ago
- 基于TrOCR + UniMER-1M数据集,训练一个小而美的公式识别模型☆30Mar 17, 2026Updated 5 months ago
- My personal implementation of SVTR model for handwritten OCR☆14Mar 1, 2024Updated 2 years ago
- The source codes of TDv2 in paper: TDv2: A Novel Tree-Structured Decoder for Offline Mathematical Expression Recognition.☆12Jul 28, 2022Updated 4 years ago
- a math-formula image recognition project which placed at the first place in a competition hosted by NAVER CONNECT boostcamp AI Tech☆10Dec 16, 2023Updated 2 years ago
- Distorted Document Images dataset (DDI-100).☆147Nov 1, 2022Updated 3 years ago
- The official code for “Deep Unrestricted Document Image Rectification”, TMM, 2023.☆537Feb 1, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A large scale camera-taken table detection and recognition dataset.☆151Apr 9, 2026Updated 4 months ago
- Faster Arbitrarily-Shaped Text Detector with Minimalist Kernel Representation☆207May 23, 2025Updated last year
- UniMERNet: A Universal Network for Real-World Mathematical Expression Recognition☆495Sep 28, 2025Updated 10 months ago
- Official Implementation of SynthTIGER (Synthetic Text Image Generator), ICDAR 2021☆579Jun 14, 2024Updated 2 years ago
- ☆189Feb 27, 2024Updated 2 years ago
- Official PyTorch implementation of "CBNet: A Plug-and-Play Network for Segmentation-Based Scene Text Detection"☆23Mar 30, 2024Updated 2 years ago
- This is the official implementation to the EMNLP 2024 paper: Modeling Layout Reading Order as Ordering Relations for Visually-rich Docume…☆32Jan 19, 2026Updated 6 months ago
- NAF-DPM: A Nonlinear Activation-Free Diffusion Probabilistic Model for Document Enhancement☆54Aug 5, 2024Updated 2 years ago
- Implementation of DocFormer: End-to-End Transformer for Document Understanding, a multi-modal transformer based architecture for the task…☆290Feb 13, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Code and data for the paper: DTSM: Toward Dense Table Structure Recognition with Text Query Encoder and Adjacent Feature Aggregator☆14Apr 28, 2024Updated 2 years ago
- [ICDAR 2023] (Oral) An End-to-End Unified Domain Adaptive Transformer for Document Instance Segmentation☆73Sep 12, 2024Updated last year
- A Faster LayoutReader Model based on LayoutLMv3, Sort OCR bboxes to reading order.☆324Aug 15, 2025Updated last year
- A collection of original, innovative ideas and algorithms towards Advanced Literate Machinery. This project is maintained by the OCR Team…☆1,834Mar 17, 2026Updated 5 months ago
- Code for CVPR21 paper A Multiplexed Network for End-to-End, Multilingual OCR☆80Dec 2, 2022Updated 3 years ago
- [MM'2024] PEneo, an effective algorithm for key-value pair extraction from form-like documents, designed for real-world applications.☆41Apr 7, 2025Updated last year
- ☆105Aug 22, 2024Updated last year
- SPRINT: Script-agnostic Structure Recognition in Tables☆17Mar 26, 2025Updated last year
- Deep learning, Convolutional neural networks, Image processing, Document processing, Table detection, Page object detection, Table classi…☆67Feb 24, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Free Persian Word Level OCR Dataset☆24Aug 1, 2020Updated 6 years ago
- layoutlmv3 在中文文档上的应用☆21May 17, 2023Updated 3 years ago
- [ICCV2023] Self-supervised Character-to-Character Distillation for Text Recognition☆153Jul 12, 2026Updated last month
- Syntax-Aware Network for Handwritten Mathematical Expression Recognition☆103Feb 21, 2023Updated 3 years ago
- Implementation of the paper: Going Full-TILT Boogie on Document Understanding with Text-Image-Layout Transformer.☆18Apr 23, 2023Updated 3 years ago
- Official code for DocNLC: A Document Image Enhancement Framework with Normalized and Latent Contrastive Representation for Multiple Degra…☆44Mar 20, 2026Updated 4 months ago
- Basic HTR concepts/modules to boost performance☆42Nov 30, 2024Updated last year