Synthetic Document Generator for Document AI. Creates document images annotated with text and bounding boxes of each word. Images contain headings, tables, paragraphs with different formatting and fonts. Can be used in OCR, document transformers pretraining, text detection and more other tasks.
☆33Jul 23, 2025Updated last year
Alternatives and similar repositories for DocumentGenerator_DoGe
Users that are interested in DocumentGenerator_DoGe are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Create realistic looking handwritten text PDFs from text files.☆15Jun 19, 2021Updated 5 years ago
- Handwritten Text Generation☆17Oct 17, 2022Updated 3 years ago
- Aggregation framework for annotating datasets in computer vision tasks (detection, segmentation, video captioning etc.)☆12Nov 6, 2024Updated last year
- This repo contains a curated list of research papers and resources focusing on Handwritten Text Generation (HTG)☆25Jan 20, 2026Updated 6 months ago
- An easy-to-run OCR model pipeline based on CRNN and CTC loss☆49Sep 25, 2025Updated 10 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Augmentation pipeline for rendering synthetic paper printing, faxing, scanning and copy machine processes☆561Jul 20, 2025Updated last year
- this repo attemps to reproduce DSOD: Learning Deeply Supervised Object Detectors from Scratch use gluon reimplementation☆14Aug 18, 2018Updated 7 years ago
- ☆12Jul 19, 2019Updated 7 years ago
- Implementation of paper - RepVGG-GELAN: ENHANCED GELAN WITH VGG-STYLE CONVNETS FOR BRAIN TUMOR DETECTION☆10Jul 19, 2025Updated last year
- PyTorch Implementation for CS229 Course Project - "Grammatical Error Correction using Neural Networks"☆10Dec 16, 2017Updated 8 years ago
- official implementation of Training-free Boost for Open-Vocabulary Object Detection with Confidence Aggregation☆13Apr 15, 2024Updated 2 years ago
- ☆11Jul 26, 2024Updated 2 years ago
- 한글 인식기☆13May 7, 2022Updated 4 years ago
- Use the MobileNet V2 as the basenet instead of the original VGG16☆14Aug 28, 2019Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- MathNet: A Data-Centric Approach, Dataset and Benchmark Model to Advance Mathematical Expression Recognition☆10Mar 19, 2025Updated last year
- SPLINE-Net: Sparse Photometric Stereo through Lighting Interpolation and Normal Estimation Networks☆11Apr 13, 2023Updated 3 years ago
- Graph-based Document Structure Analysis☆18Mar 26, 2025Updated last year
- ONNXモデルをpyca/cryptographyを用いて暗号化/復号化するサンプル☆16Mar 19, 2022Updated 4 years ago
- [MICCAI 2024] RadiomicsFill-Mammo: Synthetic Mammogram Mass Manipulation with Radiomics Features☆10Aug 22, 2025Updated 11 months ago
- ControlNet with Txt2Img | Img2Img | + Multiple LoRAs, All in one jupyter notebook for Flux.1 dev. Able to run on Google Colab Free Tier☆22Dec 1, 2024Updated last year
- ☆12Aug 19, 2023Updated 2 years ago
- Original PyTorch Implementation for the EMNLP 2023 Paper "Beyond Detection: A Defend-and-Summarize Strategy for Robust and Interpretable …☆16Dec 14, 2023Updated 2 years ago
- Estimating normal map and depth map using Frankot Chellappa algorithm in Python☆13Nov 12, 2022Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆15Jan 2, 2025Updated last year
- ☆13Jun 16, 2021Updated 5 years ago
- QTSeg: A Query Token-Based Architecture for Efficient 2D Medical Image Segmentation☆12Feb 28, 2025Updated last year
- Optimize QWen1.5 models with TensorRT-LLM☆17May 14, 2024Updated 2 years ago
- ☆21May 22, 2023Updated 3 years ago
- ACL 2025: Synthetic data generation pipelines for text-rich images.☆168Mar 1, 2025Updated last year
- Eden Flux LoRA trainer and full-finetuning☆23Mar 21, 2025Updated last year
- Official repository of "Deep Image Composition Meets Image Forgery"☆13May 30, 2024Updated 2 years ago
- The official code for “Geometric Representation Learning for Document Image Rectification”, ECCV, 2022.☆94Jun 18, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICLR 2026] DiffInk: Glyph- and Style-Aware Latent Diffusion Transformer for Text to Online Handwriting Generation☆39Jun 19, 2026Updated last month
- [ICRA 2024] GelRoller: A Rolling Vision-based Tactile Sensor for Large Surface Reconstruction Using Self-Supervised Photometric Stereo Me…☆13Sep 23, 2024Updated last year
- Code for the CubeRefine R-CNN model of our CVPRW '23 paper "Parcel3D: Shape Reconstruction From Single RGB Images for Applications in Tra…☆17Jul 12, 2023Updated 3 years ago
- This is an accurate implementation for IoU loss between two rotated polygons. This algorithm is accurate and differential, but there is n…☆18Mar 5, 2022Updated 4 years ago
- Generates a ready to submit arxiv zip file out of your paper.☆11Feb 15, 2020Updated 6 years ago
- ☆13Aug 22, 2024Updated last year
- Make writing easier!☆83Aug 15, 2021Updated 4 years ago