[ICDAR 2023] SelfDocSeg: A self-supervised vision-based approach towards Document Segmentation (Oral)
☆43Oct 6, 2023Updated 2 years ago
Alternatives and similar repositories for SelfDocSeg
Users that are interested in SelfDocSeg are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆43Jun 15, 2024Updated 2 years ago
- A Bottom-Up Instance Segmentation Strategy for segmenting document instances using Transformers☆59Sep 9, 2024Updated last year
- NeurIPS'23 & AAAI'24 Workshop Oral - High-Performance Transformers for Table Structure Recognition Need Early Convolutions☆45Apr 21, 2026Updated 4 months ago
- TRACE: Table Reconstruction Aligned to Corner and Edges (ICDAR 2023)☆32Mar 13, 2024Updated 2 years ago
- ☆12Mar 20, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- The code and data for "Summary-Oriented Vision Modeling for Multimodal Abstractive Summarization"☆11May 16, 2023Updated 3 years ago
- [ICDAR 2024] (Best Student Paper🏆) Exploring Knowledge Distillation Towards Document Object Detection with Structured Graph Creation☆14Sep 6, 2024Updated last year
- An official implementation of paper "Paragraph2Graph: A Language-independent GNN-based framework for layout analysis"☆82Oct 14, 2023Updated 2 years ago
- Table detection (TD) and table structure recognition (TSR) using Yolov5/Yolov8, and you can get the same (even better) result compared wi…☆52Jul 3, 2024Updated 2 years ago
- [WACV 2026 Round 1] Beyond Single Object Text-to-SVG Synthesis with Comprehensive Canvas Layout☆22Oct 11, 2025Updated 10 months ago
- 利用Swin-Unet(Swin Transformer Unet)实现对文档图片里表格结构的识别,Swin-unet (Swin Transformer Unet) is used to identify the document table structure☆27Feb 23, 2024Updated 2 years ago
- Official Repository of "Fibottention: Inceptive Visual Representation Learning with Diverse Attention Across Heads"☆17Aug 24, 2026Updated last week
- TAT-DQA: Towards Complex Document Understanding By Discrete Reasoning☆25Sep 17, 2024Updated last year
- ICDAR 2024/2026 Table OCR Model☆39Jun 16, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- The official code for "OG-HFYOLO :Orientation Gradient Guidance and Heterogeneous Feature Fusion For Deformation Table Cell Instance Segm…☆13Jul 28, 2025Updated last year
- Table Structure Recognition☆83Mar 11, 2023Updated 3 years ago
- time-series row column classification☆14Jan 7, 2022Updated 4 years ago
- MTL-TabNet: Multi-task Learning based Model for Image-based Table Recognition☆103May 30, 2024Updated 2 years ago
- This is an official implementation for the WTW Dataset in "Parsing Table Structures in the Wild " on table detection and table structure …☆183Sep 15, 2021Updated 4 years ago
- ☆46Feb 7, 2023Updated 3 years ago
- RoDLA: Benchmarking the Robustness of Document Layout Analysis Models☆39Mar 26, 2025Updated last year
- [MM'2024] PEneo, an effective algorithm for key-value pair extraction from form-like documents, designed for real-world applications.☆41Apr 7, 2025Updated last year
- (CVPR 2026) TRivia: Self-supervised Fine-tuning of Vision-Language Models for Table Recognition☆35Jul 14, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A curated list of resources dedicated to table recognition☆405Dec 12, 2024Updated last year
- Ongoing research training transformer language models at scale, including: BERT & GPT-2☆19Jul 20, 2023Updated 3 years ago
- Dataset of PNG images from synthetically generated table layouts with annotations in JSONL files☆154Sep 17, 2025Updated 11 months ago
- Datasets and Evaluation Scripts for CompHRDoc☆59Feb 25, 2025Updated last year
- 该项目是为了使用layoutlmv3针对中文图片训练和推理。 其中主要解决三个问题: 1.数据标准化成可以的训练数据集格式 2.layoutlmv3-base-chinese 分词修改 2.超过512长度的文本切分和滑窗操作☆65Sep 6, 2024Updated last year
- [WACV2023] This is the official PyTorch impelementation of our paper "[Rethinking Rotation in Self-Supervised Contrastive Learning: Adapt…☆12Feb 24, 2023Updated 3 years ago
- an unofficial code for augment-XY-CUT in XYLayoutLM☆30Jul 12, 2022Updated 4 years ago
- [ECCV 2022] Learning Instance-Specific Adaptation for Cross-Domain Segmentation☆14Jul 17, 2022Updated 4 years ago
- Code and data for the paper: DTSM: Toward Dense Table Structure Recognition with Text Query Encoder and Adjacent Feature Aggregator☆14Apr 28, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [ICML 2022 Spotlight] Finding the Task-Optimal Low-Bit Sub-Distribution in Deep Neural Networks☆11May 21, 2023Updated 3 years ago
- Kompakkt - the web based 3D viewer and 3D annotation system.☆17Aug 17, 2026Updated 2 weeks ago
- Repo☆13Mar 7, 2022Updated 4 years ago
- Research project on the state of the field of Multilingual Digital Humanities, with an initial focus on Arabic☆13Aug 20, 2026Updated last week
- The repository provides code for training the SegmentAnything Model (SAM) for predicting frame polygons in comic books☆58Mar 14, 2024Updated 2 years ago
- Image to Latex using Encoder-Decoder architecture☆19May 21, 2025Updated last year
- Official PyTorch Implementation of "WordStylist: Styled Verbatim Handwritten Text Generation with Latent Diffusion Models" - ICDAR 2023☆82Jun 25, 2024Updated 2 years ago