[CVPR 2023 highlight] Towards Flexible Multi-modal Document Models
☆59Sep 7, 2023Updated 3 years ago
Alternatives and similar repositories for flex-dm
Users that are interested in flex-dm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of CanvasVAE: Learning to Generate Vector Graphic Documents, ICCV 2021☆71Mar 7, 2023Updated 3 years ago
- [CVPR 2023] LayoutDM: Discrete Diffusion Model for Controllable Layout Generation☆297Oct 24, 2023Updated 2 years ago
- Official repository for "PosterLayout: A New Benchmark and Approach for Content-aware Visual-Textual Presentation Layout" (CVPR 2023).☆149Mar 31, 2025Updated last year
- Official implementation of Generative Colorization of Structured Mobile Web Pages, WACV 2023.☆21Dec 7, 2023Updated 2 years ago
- ☆81Feb 14, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆153Jan 31, 2024Updated 2 years ago
- Layout Generation and Baseline implementations☆167Jul 10, 2022Updated 4 years ago
- An awesome list of layout generation papers☆275Mar 22, 2025Updated last year
- This is a data repository for the ACL 2020 paper: "Let Me Choose: From Verbal Context to Font Selection"☆11May 5, 2020Updated 6 years ago
- Code for the paper "Harmonious Textual Layout Generation over Natural Images via Deep Aesthetics Learning" (TMM 2021)☆41Aug 3, 2022Updated 4 years ago
- PyTorch implementation of "LayoutTransformer: Layout Generation and Completion with Self-attention" to appear in ICCV 2021☆169Jan 25, 2022Updated 4 years ago
- Official Repo of Graphist☆133Apr 23, 2024Updated 2 years ago
- OCR-VQGAN, a discrete image encoder (tokenizer and detokenizer) for figure images in Paper2Fig100k dataset. Implementation of OCR Percept…☆86Jan 30, 2023Updated 3 years ago
- This is a repository for the ACL 2020 paper: "Let Me Choose: From Verbal Context to Font Selection"☆12Nov 21, 2022Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [MM2023] An official implement of the paper "One-stage Low-resolution Text Recognition with High-resolution Knowledge Transfer"☆16Nov 3, 2023Updated 2 years ago
- Cheng-Fu Yang*, Wan-Cyuan Fan*, Fu-En Yang, Yu-Chiang Frank Wang, "LayoutTransformer: Scene Layout Generation with Conceptual and Spatial…☆64Apr 3, 2022Updated 4 years ago
- Official code for paper: Desigen: A Pipeline for Controllable Design Template Generation [CVPR'24]☆75Jul 18, 2024Updated 2 years ago
- ☆211Jan 6, 2025Updated last year
- The official PyTorch implementation for arXiv'23 paper 'LayoutDETR: Detection Transformer Is a Good Multimodal Layout Designer'☆108Jul 24, 2026Updated 2 months ago
- [CVPR 2024 Oral] Official repository for RALF: Retrieval-Augmented Layout Transformer for Content-Aware Layout Generation☆143Jul 6, 2024Updated 2 years ago
- [CVPR 2021] Rethinking Text Segmentation: A Novel Dataset and A Text-Specific Refinement Approach☆276Dec 2, 2023Updated 2 years ago
- Collection of Aesthetics Assessment Papers for Graphic Designs.☆47Aug 11, 2026Updated last month
- This is the official repository for "Can GPTs Evaluate Graphic Design Based on Design Principles?".☆16Feb 10, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Resources about Design + AI (papers, datasets, events, companies, etc.)☆65Mar 17, 2021Updated 5 years ago
- ☆26Oct 20, 2022Updated 3 years ago
- ☆28May 22, 2021Updated 5 years ago
- Official repo for NeurIPS 2023 paper "LayoutGPT: Compositional Visual Planning and Generation with Large Language Models"☆408Apr 10, 2024Updated 2 years ago
- Official implementation of the MM'21 paper "Constrained Graphic Layout Generation via Latent Optimization" (LayoutGAN++, CLG-LO, and Layo…☆139Jul 24, 2023Updated 3 years ago
- This repository is the implementation of "Don't Forget Me: Accurate Background Recovery for Text Removal via Modeling Local-Global Contex…☆97Feb 21, 2023Updated 3 years ago
- Text-To-Image Generation with Chinese Characters☆132Jul 20, 2023Updated 3 years ago
- DreamDance: Personalized Text-to-video Generation by Combining Text-to-Image Synthesis and Motion Transfer☆14Dec 16, 2022Updated 3 years ago
- ☆30Sep 12, 2022Updated 4 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- This repository is the code of our paper "DiffUTE: Universal Text Editing Diffusion Model" (NeurIPS'2023).☆144Apr 11, 2025Updated last year
- Project page for "MG-Gen: Single Image to Motion Graphics Generation with Layer Decomposition"☆17Apr 18, 2025Updated last year
- ☆43Sep 12, 2024Updated 2 years ago
- MetricEval: A framework that conceptualizes and operationalizes four main components of metric evaluation, in terms of reliability and va…☆13Nov 6, 2023Updated 2 years ago
- STIRER: A Unified Model for Low-Resolution Scene Text Image Recovery and Recognition -- ACMMM 2023☆14Dec 2, 2024Updated last year
- Official release of the benchmark in paper "VSP: Diagnosing the Dual Challenges of Perception and Reasoning in Spatial Planning Tasks for…☆23Aug 1, 2025Updated last year
- Real-CE: A Benchmark for Chinese-English Scene Text Image Super-resolution (ICCV2023)☆100Nov 3, 2023Updated 2 years ago