☆48Feb 7, 2025Updated last year
Alternatives and similar repositories for Ocean-OCR
Users that are interested in Ocean-OCR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆18Jul 24, 2025Updated last year
- ☆40Oct 7, 2023Updated 2 years ago
- SPRINT: Script-agnostic Structure Recognition in Tables☆18Mar 26, 2025Updated last year
- ☆197Dec 7, 2025Updated 9 months ago
- [ICCV 2023] Code base for Revisiting Scene Text Recognition: A Data Perspective☆204Nov 1, 2023Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- official code for "Fox: Focus Anywhere for Fine-grained Multi-page Document Understanding"☆198May 31, 2024Updated 2 years ago
- ☆143Feb 13, 2024Updated 2 years ago
- ☆26Apr 9, 2025Updated last year
- The source code repository for the paper.☆26Sep 8, 2025Updated last year
- Handwritten Text Recognition and Character Detection☆174Sep 28, 2025Updated 11 months ago
- ☆171Sep 14, 2026Updated last week
- [AAAI'23 Oral] DPText-DETR: Towards Better Scene Text Detection with Dynamic Points in Transformer☆205Aug 31, 2023Updated 3 years ago
- [NAACL 2024] Visually Guided Generative Text-Layout Pre-training for Document Intelligence☆148Sep 10, 2024Updated 2 years ago
- Table Structure Recognition☆83Mar 11, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆42Sep 2, 2023Updated 3 years ago
- On the Hidden Mystery of OCR in Large Multimodal Models (OCRBench)☆894Updated this week
- ☆18Jul 9, 2024Updated 2 years ago
- [MM'2024] Official release of RFUND introduced in the MM'2024 paper "PEneo: Unifying Line Extraction, Line Grouping, and Entity Linking f…☆21Dec 4, 2024Updated last year
- DatasetImgLabeler is a image annotation tool for researchers to prepare datasets in ICDAR2015 format☆12Dec 7, 2019Updated 6 years ago
- ☆15Jul 11, 2022Updated 4 years ago
- A paper collection of recent diffusion models for text-image generation tasks, e,g., visual text generation, font generation, text remova…☆272Dec 19, 2024Updated last year
- A full codebase for replicating the results of Nougat from downloading arXiv dataset to the final evaluation. It also contains a few fixe…☆11Dec 11, 2023Updated 2 years ago
- (ACL 2025) MAmmoTH-VL: Eliciting Multimodal Reasoning with Instruction Tuning at Scale☆50Jun 4, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A Unified Framework for Document Parsing Tasks (Including Document Layout Analysis, OCR, Formula Recognition, and Table Recognition)☆15Jul 1, 2025Updated last year
- A zero-shot faithfulness evaluation metric for text summarization☆11Oct 17, 2023Updated 2 years ago
- The source codes of TDv2 in paper: TDv2: A Novel Tree-Structured Decoder for Offline Mathematical Expression Recognition.☆12Jul 28, 2022Updated 4 years ago
- Table Structure Recognition☆28Jul 25, 2024Updated 2 years ago
- An unofficial Github Copilot extension for Lapce☆15Mar 12, 2024Updated 2 years ago
- The proposed simulated dataset consisting of 9,536 charts and associated data annotations in CSV format.☆26Feb 22, 2024Updated 2 years ago
- ☆13Jun 10, 2025Updated last year
- Project code for ACM MM2020 paper: "TextRay: Contour-based Geometric Modeling for Arbitrary-shaped Scene Text Detection"☆47Oct 3, 2023Updated 2 years ago
- INF Tech's open-source MLLMs for SOTA visual-language understanding and advanced document intelligence.☆255Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICASSP2024] An official implement of the paper "EFFICIENT SCENE TEXT IMAGE SUPER-RESOLUTION WITH SEMANTIC GUIDANCE"☆25May 12, 2024Updated 2 years ago
- We identify the desiderata for a comprehensive benchmark and propose Visually Rich Document Understanding (VRDU). VRDU contains two datas…☆83Feb 8, 2023Updated 3 years ago
- Opam2 remote for beta versions of the OCaml compiler☆16May 1, 2019Updated 7 years ago
- [IJCAI-2024] The official code of Self-Supervised Pre-training with Symmetric Superimposition Modeling for Scene Text Recognition☆10Aug 10, 2025Updated last year
- ACL'2024-Main: Synergetic Event Understanding: A Collaborative Approach to Cross-Document Event Coreference Resolution with Large Languag…☆12Sep 19, 2025Updated last year
- What Is a Good Caption? A Comprehensive Visual Caption Benchmark for Evaluating Both Correctness and Thoroughness☆28May 16, 2025Updated last year
- [ACL 2025 main] The official GitHub page of "Reviving Cultural Heritage: A Novel Approach for Comprehensive Historical Document Restorati…☆65Jun 28, 2026Updated 3 months ago