[CVPR 2023 highlight] Towards Flexible Multi-modal Document Models
☆59Sep 7, 2023Updated 2 years ago
Alternatives and similar repositories for flex-dm
Users that are interested in flex-dm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of CanvasVAE: Learning to Generate Vector Graphic Documents, ICCV 2021☆71Mar 7, 2023Updated 3 years ago
- [CVPR 2023] LayoutDM: Discrete Diffusion Model for Controllable Layout Generation☆300Oct 24, 2023Updated 2 years ago
- Official repository for "PosterLayout: A New Benchmark and Approach for Content-aware Visual-Textual Presentation Layout" (CVPR 2023).☆148Mar 31, 2025Updated last year
- Official implementation of Generative Colorization of Structured Mobile Web Pages, WACV 2023.☆22Dec 7, 2023Updated 2 years ago
- ☆82Feb 14, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆153Jan 31, 2024Updated 2 years ago
- Layout Generation and Baseline implementations☆167Jul 10, 2022Updated 4 years ago
- An awesome list of layout generation papers☆273Mar 22, 2025Updated last year
- This is a data repository for the ACL 2020 paper: "Let Me Choose: From Verbal Context to Font Selection"☆11May 5, 2020Updated 6 years ago
- Code for the paper "Harmonious Textual Layout Generation over Natural Images via Deep Aesthetics Learning" (TMM 2021)☆41Aug 3, 2022Updated 4 years ago
- PyTorch implementation of "LayoutTransformer: Layout Generation and Completion with Self-attention" to appear in ICCV 2021☆169Jan 25, 2022Updated 4 years ago
- OCR-VQGAN, a discrete image encoder (tokenizer and detokenizer) for figure images in Paper2Fig100k dataset. Implementation of OCR Percept…☆85Jan 30, 2023Updated 3 years ago
- This is a repository for the ACL 2020 paper: "Let Me Choose: From Verbal Context to Font Selection"☆12Nov 21, 2022Updated 3 years ago
- [MM2023] An official implement of the paper "One-stage Low-resolution Text Recognition with High-resolution Knowledge Transfer"☆16Nov 3, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Cheng-Fu Yang*, Wan-Cyuan Fan*, Fu-En Yang, Yu-Chiang Frank Wang, "LayoutTransformer: Scene Layout Generation with Conceptual and Spatial…☆64Apr 3, 2022Updated 4 years ago
- Official code for paper: Desigen: A Pipeline for Controllable Design Template Generation [CVPR'24]☆75Jul 18, 2024Updated 2 years ago
- ☆209Jan 6, 2025Updated last year
- The official PyTorch implementation for arXiv'23 paper 'LayoutDETR: Detection Transformer Is a Good Multimodal Layout Designer'☆107Jul 24, 2026Updated 2 weeks ago
- ☆168Aug 1, 2026Updated last week
- [CVPR 2024 Oral] Official repository for RALF: Retrieval-Augmented Layout Transformer for Content-Aware Layout Generation☆143Jul 6, 2024Updated 2 years ago
- [CVPR 2021] Rethinking Text Segmentation: A Novel Dataset and A Text-Specific Refinement Approach☆275Dec 2, 2023Updated 2 years ago
- This is the official repository for "Can GPTs Evaluate Graphic Design Based on Design Principles?".☆13Feb 10, 2025Updated last year
- Resources about Design + AI (papers, datasets, events, companies, etc.)☆64Mar 17, 2021Updated 5 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆26Oct 20, 2022Updated 3 years ago
- ☆28May 22, 2021Updated 5 years ago
- Official implementation of the MM'21 paper "Constrained Graphic Layout Generation via Latent Optimization" (LayoutGAN++, CLG-LO, and Layo…☆139Jul 24, 2023Updated 3 years ago
- This repository is the implementation of "Don't Forget Me: Accurate Background Recovery for Text Removal via Modeling Local-Global Contex…☆97Feb 21, 2023Updated 3 years ago
- Text-To-Image Generation with Chinese Characters☆133Jul 20, 2023Updated 3 years ago
- BTS: A Bi-lingual Benchmark for Text Segmentation in the Wild☆33Apr 16, 2024Updated 2 years ago
- DreamDance: Personalized Text-to-video Generation by Combining Text-to-Image Synthesis and Motion Transfer☆14Dec 16, 2022Updated 3 years ago
- ☆30Sep 12, 2022Updated 3 years ago
- This repository is the code of our paper "DiffUTE: Universal Text Editing Diffusion Model" (NeurIPS'2023).☆143Apr 11, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Project page for "MG-Gen: Single Image to Motion Graphics Generation with Layer Decomposition"☆16Apr 18, 2025Updated last year
- ☆43Sep 12, 2024Updated last year
- MetricEval: A framework that conceptualizes and operationalizes four main components of metric evaluation, in terms of reliability and va…☆12Nov 6, 2023Updated 2 years ago
- ☆10Nov 22, 2022Updated 3 years ago
- Official release of the benchmark in paper "VSP: Diagnosing the Dual Challenges of Perception and Reasoning in Spatial Planning Tasks for…☆22Aug 1, 2025Updated last year
- Real-CE: A Benchmark for Chinese-English Scene Text Image Super-resolution (ICCV2023)☆99Nov 3, 2023Updated 2 years ago
- ☆99Jan 3, 2024Updated 2 years ago