【2024 ECAI】First Creating Backgrounds Then Rendering Texts: A New Paradigm for Visual Text Blending
☆14Jun 16, 2025Updated last year
Alternatives and similar repositories for GlyphOnly
Users that are interested in GlyphOnly are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The official project of paper "Visual Text Processing: A Comprehensive Review and Unified Evaluation""☆104Oct 20, 2025Updated 10 months ago
- [2024-NeurIPS] TextCtrl: Diffusion-based Scene Text Editing with Prior Guidance Control☆107Mar 16, 2025Updated last year
- Code Implementation of the Paper: EasyText: Controllable Diffusion Transformer for Multilingual Text Rendering☆58Jun 16, 2025Updated last year
- Code for "Translatotron-V(ison): An End-to-End Model for In-Image Machine Translation" (Findings of ACL 2024)☆16Jul 4, 2024Updated 2 years ago
- Linguistics-aware Masked Image Modeling for Self-supervised Scene Text Recognition☆15Oct 26, 2025Updated 10 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Optocal Character Recognition (OCR / HTR) using Transformers☆11Aug 20, 2022Updated 4 years ago
- ControlText: Unlocking Controllable Fonts in Multilingual Text Rendering without Font Annotations☆36Apr 3, 2025Updated last year
- TextFlux: An OCR-Free DiT Model for High-Fidelity Multilingual Scene Text Synthesis☆100Sep 18, 2025Updated 11 months ago
- Project page of "2026-ICLR Echo: Towards Advanced Audio Comprehension via Audio-Interleaved Reasoning"☆17Mar 26, 2026Updated 5 months ago
- 激光SLAM理论与实践课程学习记录☆12Apr 22, 2022Updated 4 years ago
- [ICCV 2025] FiVE-Bench: A Fine-grained Video Editing Benchmark for Evaluating Emerging Diffusion and Rectified Flow Models☆19Aug 26, 2025Updated last year
- ☆98Jan 3, 2024Updated 2 years ago
- Fleming-VL: Towards Universal Medical Visual Understanding with Multimodal LLMs☆15Nov 6, 2025Updated 9 months ago
- TextCrafter: Accurately Rendering Multiple Texts in Complex Visual Scenes☆97Nov 26, 2025Updated 9 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆24May 11, 2024Updated 2 years ago
- Code for Decomposed Vector-Quantized Variational Autoencoder for Human Grasp Generation☆25Apr 15, 2025Updated last year
- TMI 2023: Less is More: Surgical Phase Recognition from Timestamp Supervision☆23Feb 9, 2023Updated 3 years ago
- ☆11Sep 10, 2024Updated last year
- Code and data for the paper: DTSM: Toward Dense Table Structure Recognition with Text Query Encoder and Adjacent Feature Aggregator☆14Apr 28, 2024Updated 2 years ago
- 1.0☆15Jun 7, 2025Updated last year
- ☆15Nov 26, 2023Updated 2 years ago
- DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning☆18Jun 14, 2026Updated 2 months ago
- An official Project related to Paper "Perceiving Ambiguity and Semantics without Recognition: An Efficient and Effective Ambiguous Scene …☆22Dec 3, 2023Updated 2 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- 中国科学院大学研究生课程 模式识别与机器学习☆16Jan 8, 2022Updated 4 years ago
- ACL'2024-Main: Synergetic Event Understanding: A Collaborative Approach to Cross-Document Event Coreference Resolution with Large Languag…☆12Sep 19, 2025Updated 11 months ago
- 【ICDAR 2024】Coarse-to-Fine Document Image Registration for Dewarping☆24Jul 15, 2024Updated 2 years ago
- ☆28Nov 29, 2023Updated 2 years ago
- Turn every moment into momentum☆22Jun 1, 2026Updated 2 months ago
- Pytorch implementation for the pilot study on the robustness of latent diffusion models.☆13Jun 20, 2023Updated 3 years ago
- resources for text detection, text recognition, and end to end text spotting☆13Apr 23, 2023Updated 3 years ago
- Code of the paper Unsupervised Domain Adaptation through Shape Modeling for Medical Image Segmentation.☆16Oct 16, 2022Updated 3 years ago
- [CVPR2026] TextPecker: Rewarding Structural Anomaly Quantification for Enhancing Visual Text Rendering☆58Jul 17, 2026Updated last month
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Structured Domain Adaptation with Online Relation Regularization for Unsupervised Person Re-ID☆18Jun 9, 2020Updated 6 years ago
- Official code repo of Video-Browser: Towards Agentic Open-web Video Browsing☆28Jan 19, 2026Updated 7 months ago
- Text-To-Image Generation with Chinese Characters☆133Jul 20, 2023Updated 3 years ago
- The official code of Linguistic More: Taking a Further Step toward Efficient and Accurate Scene Text Recognition (IJCAI2023)☆26Sep 3, 2023Updated 2 years ago
- RepText: Rendering Visual Text via Replicating 🔥☆138Jun 7, 2025Updated last year
- 2022高教社杯数学建模 C题 古代玻璃制品的成分分析与鉴别☆14Nov 26, 2022Updated 3 years ago
- Official implementation of VLPCook: Vision and Structured-Language Pretraining for Cross-Modal Food Retrieval☆16Mar 25, 2023Updated 3 years ago