Comics Dataset Framework for Comics Understanding
☆44Sep 1, 2025Updated last year
Alternatives and similar repositories for CoMix
Users that are interested in CoMix are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Repository for "CoMix: Comprehensive Benchmark for Multi-Task Comic Understanding"☆18Nov 20, 2024Updated last year
- [ECCV-W] Official repo for the paper "ComiCap: A VLMs pipeline for dense captioning of Comic Panels"☆15Nov 20, 2024Updated last year
- Original Full Repository of the Paper: "Domain-Adaptive Self-Supervised Pre-training for Face & Body Detection in Drawings"☆20Oct 14, 2025Updated 11 months ago
- MangaLMM – Try the official demo below☆49Nov 9, 2025Updated 10 months ago
- Official PyTorch Implementation of "Rethinking HTG Evaluation: Bridging Generation and Recognition" (Oral) - 1st Workshop on Critical Eva…☆17Sep 23, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Searching a High Performance Feature Extractor for Text Recognition Network. TPAMI 2022☆13Nov 25, 2022Updated 3 years ago
- Official repository of Manga109Dialog (ICME 2024)☆30Aug 3, 2024Updated 2 years ago
- This repo contains a curated list of research papers and resources focusing on Handwritten Text Generation (HTG)☆27Jan 20, 2026Updated 8 months ago
- A framework for efficient model inference with omni-modality models☆30Sep 3, 2026Updated 3 weeks ago
- TextAdaIN: Paying Attention to Shortcut Learning in Text Recognizers☆21Jul 26, 2022Updated 4 years ago
- There's no publicly available free-to-use Manga Dataset, so I decided to make one artificially!☆55Sep 2, 2023Updated 3 years ago
- baselines for DocVQA dataset☆21Apr 11, 2021Updated 5 years ago
- Basic HTR concepts/modules to boost performance☆42Nov 30, 2024Updated last year
- OCR-VQGAN, a discrete image encoder (tokenizer and detokenizer) for figure images in Paper2Fig100k dataset. Implementation of OCR Percept…☆86Jan 30, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- List of diffusion papers accepted in ECCV 2024.☆15Oct 17, 2024Updated last year
- (Unstructured) Weight Pruning via Adaptive Sparsity Loss☆15Sep 28, 2022Updated 3 years ago
- The official repo of the CVPR 2026 paper UniLight☆20Sep 1, 2026Updated 3 weeks ago
- [ICDAR 2024] (Best Student Paper🏆) Exploring Knowledge Distillation Towards Document Object Detection with Structured Graph Creation☆15Sep 6, 2024Updated 2 years ago
- Facial Alignment for Anime Styled Faces☆10Mar 26, 2021Updated 5 years ago
- Domain Adaptation for anime face detection☆14Nov 25, 2019Updated 6 years ago
- [ICCV 2025] What Changed? Detecting and Evaluating Instruction-Guided Image Edits with Multimodal Large Language Models☆16Nov 3, 2025Updated 10 months ago
- [AAAI 2025 (Oral)] SAIL: Sample-Centric In-Context Learning for Document Information Extraction☆19Dec 24, 2024Updated last year
- [CVPR 2024 Highlight] - Stationary Representations: Optimally Approximating Compatibility and Implications for Improved Model Replacement…☆14Oct 21, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Official repository of the paper: "A Comprehensive Gold Standard and Benchmark for Comics Text Detection and Recognition"☆27Jul 10, 2023Updated 3 years ago
- OCR Annotations from Amazon Textract for Industry Documents Library☆105Aug 20, 2022Updated 4 years ago
- ☆27Feb 20, 2024Updated 2 years ago
- LLM frontend for roleplay☆78Sep 17, 2026Updated last week
- ☆36Jun 22, 2023Updated 3 years ago
- ☆15Jul 31, 2020Updated 6 years ago
- Segmentation of text in manga images☆142Feb 6, 2021Updated 5 years ago
- PyTorch implementation of BMVC2022 paper Masked Vision-Language Transformers for Scene Text Recognition☆28Nov 11, 2022Updated 3 years ago
- ☆12Sep 15, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Implementation of the paper: "BRAVE : Broadening the visual encoding of vision-language models"☆26Jun 22, 2026Updated 3 months ago
- Solve the berth allocation problem using genetic-algorithm.☆11Jun 8, 2017Updated 9 years ago
- Magnification Prior: A Self-Supervised Method for Learning Representations on Breast Cancer Histopathological Images (WACV 2023)☆15Mar 13, 2023Updated 3 years ago
- Text-DIAE: A Self-Supervised Degradation Invariant Autoencoders for Text Recognition and Document Enhancement - AAAI 2023☆30Jul 12, 2023Updated 3 years ago
- Mini Model Daemon☆13Nov 9, 2024Updated last year
- WebUI extension for InteractDiffusion☆18Mar 11, 2024Updated 2 years ago
- A collection of AWESOME things about domain adaptive object detection.☆14Apr 15, 2021Updated 5 years ago