RoDLA: Benchmarking the Robustness of Document Layout Analysis Models
☆40Mar 26, 2025Updated last year
Alternatives and similar repositories for RoDLA
Users that are interested in RoDLA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆43Jun 15, 2024Updated 2 years ago
- Github repo for referring atomic video action recognition☆20Oct 2, 2024Updated last year
- ☆169Aug 31, 2026Updated last week
- Official code for DocNLC: A Document Image Enhancement Framework with Normalized and Latent Contrastive Representation for Multiple Degra…☆45Mar 20, 2026Updated 5 months ago
- Official repository for paper "Scene-agnostic Pose Regression for Visual Localization" (SPR), CVPR 2025☆34Mar 26, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆19Aug 13, 2024Updated 2 years ago
- Official repository accompaying the ICDAR 2023 paper☆14Oct 3, 2023Updated 2 years ago
- ☆16May 14, 2024Updated 2 years ago
- ☆24Jun 17, 2025Updated last year
- A Unified Framework for Document Parsing Tasks (Including Document Layout Analysis, OCR, Formula Recognition, and Table Recognition)☆15Jul 1, 2025Updated last year
- [WACV 2025] High-Fidelity Document Stain Removal via A Large-Scale Real-World Dataset and A Memory-Augmented Transformer☆24Jan 14, 2026Updated 7 months ago
- This project aims to generate syntactichandwritten mathematical expression. The dataset is generated from the CROHME 2014 training set.☆14Feb 24, 2022Updated 4 years ago
- MathNet: A Data-Centric Approach, Dataset and Benchmark Model to Advance Mathematical Expression Recognition☆10Mar 19, 2025Updated last year
- Official Repository of RefChartQA: Grounding Visual Answer on Chart Images through Instruction Tuning☆15Jul 9, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICME 2023] FlowText: Synthesizing Realistic Scene Text Video with Optical Flow Estimation☆13May 13, 2023Updated 3 years ago
- Official repository for paper "Open Panoramic Segmentation" (OPS), ECCV 2024☆38Oct 7, 2025Updated 11 months ago
- Datasets and Evaluation Scripts for CompHRDoc☆59Feb 25, 2025Updated last year
- This repository contains source codes for SoftCTC. Original paper can be found here: https://arxiv.org/abs/2212.02135☆19Mar 7, 2023Updated 3 years ago
- PyTorch implementation of BMVC2022 paper Masked Vision-Language Transformers for Scene Text Recognition☆28Nov 11, 2022Updated 3 years ago
- A curated list of resources on Document Layout Analysis☆12Aug 7, 2025Updated last year
- ☆21Mar 15, 2022Updated 4 years ago
- Graph-based Document Structure Analysis☆19Aug 1, 2026Updated last month
- Official PyTorch implementation of `[ACMMM 2023]Relational Contrastive Learning for Scene Text Recognition`☆17Sep 22, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ACM Multimedia 2023: DocDiff: Document Enhancement via Residual Diffusion Models. Also contains 1597 red seals in Chinese scenes, along w…☆356Aug 22, 2024Updated 2 years ago
- [ICDAR 2023] (Oral) An End-to-End Unified Domain Adaptive Transformer for Document Instance Segmentation☆74Sep 12, 2024Updated last year
- ☆103Aug 1, 2024Updated 2 years ago
- CTE: Contextualized Table Extraction Dataset☆17Feb 23, 2023Updated 3 years ago
- Trained Detectron2 object detection models for document layout analysis based on PubLayNet dataset☆28Apr 16, 2023Updated 3 years ago
- ☆44Jul 9, 2024Updated 2 years ago
- Implementation of the paper: Going Full-TILT Boogie on Document Understanding with Text-Image-Layout Transformer.☆18Apr 23, 2023Updated 3 years ago
- IEEE/CVF International Conference on Computer Vision Workshop (2023)☆17Feb 7, 2024Updated 2 years ago
- Evaluation of the Optical Character Recognition (OCR) capabilities of GPT-4V(ision)☆129Nov 13, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [AAAI2024] FontDiffuser: One-Shot Font Generation via Denoising Diffusion with Multi-Scale Content Aggregation and Style Contrastive Lear…☆560Mar 14, 2024Updated 2 years ago
- ☆25Jul 31, 2024Updated 2 years ago
- TAT-DQA: Towards Complex Document Understanding By Discrete Reasoning☆26Sep 17, 2024Updated last year
- Repository for the KVP10k dataset☆23Sep 18, 2025Updated 11 months ago
- [CVPR'24] Handwritten Mathematical Expressions Generation (HMEG)☆34Jun 3, 2024Updated 2 years ago
- 使用FastAPI构建发票识别系统后端服务,支持并发。使用ERFNet模型训练发票轮廓检测,进行畸变矫正,OCR识别,模板匹配,支持倾斜发票识别。准确率99.9%。☆13May 8, 2025Updated last year
- ☆16Jan 30, 2022Updated 4 years ago