The official repo of the Comics Survey: "A missing piece in Vision and Language: A Survey on Comics Understanding"
β141Jan 2, 2025Updated last year
Alternatives and similar repositories for awesome-comics-understanding
Users that are interested in awesome-comics-understanding are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- β32May 28, 2025Updated last year
- [ICDAR 2024] (Best Student Paperπ) Exploring Knowledge Distillation Towards Document Object Detection with Structured Graph Creationβ14Sep 6, 2024Updated last year
- Repository for "CoMix: Comprehensive Benchmark for Multi-Task Comic Understanding"β18Nov 20, 2024Updated last year
- [CVPR 2024 Highlight] - Stationary Representations: Optimally Approximating Compatibility and Implications for Improved Model Replacementβ¦β14Oct 21, 2024Updated last year
- Let there be clock in the beach - WACV 2022β15Nov 15, 2021Updated 4 years ago
- Open source password manager - Proton Pass β’ AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- [WACV 2026 Round 1] Beyond Single Object Text-to-SVG Synthesis with Comprehensive Canvas Layoutβ22Oct 11, 2025Updated 10 months ago
- Hadwritten Text Recognition in Few-shot Scenarioβ22Mar 25, 2023Updated 3 years ago
- This repo contains the code of "Contrastive Supervised Distillation for Continual Representation Learning", Tommaso Barletti, NiccolΓ² Bioβ¦β20Jul 5, 2022Updated 4 years ago
- Doc2Graph transforms documents into graphs and exploit a GNN to solve several tasks.β140Oct 18, 2025Updated 10 months ago
- [WACV 2024] - Reference-based Restoration of Digitized Analog Videotapesβ61Feb 11, 2024Updated 2 years ago
- [CVPR 2023] NEFER a Dataset for Neuromorphic Event-based Facial Expression Recognitionβ30Dec 7, 2023Updated 2 years ago
- [ICLR 2024] - Elastic Feature Consolidation for Cold Start Exemplar-Free Incremental Learningβ36May 26, 2025Updated last year
- [ICCV 2023] - Composed Image Retrieval on Common Objects in context (CIRCO) datasetβ87Aug 6, 2025Updated last year
- Optocal Character Recognition (OCR / HTR) using Transformersβ11Aug 20, 2022Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits β’ AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [IEEE TMM 2023] This is the official repo of the paper "Perceptual Quality Improvement in Videoconferencing using Keyframes-based GAN".β17Dec 10, 2024Updated last year
- β27Mar 7, 2025Updated last year
- Official evaluation scripts and baseline prompts for the DocVQA 2026 (ICDAR 2026) Competition on Multimodal Reasoning over Documents.β18Mar 16, 2026Updated 5 months ago
- Official PyTorch Implementation of DocSynth: A Layout Guided Approach for Controllable Document Image Synthesis - ICDAR 2021β95Jul 16, 2021Updated 5 years ago
- [ECCV'24] [TPAMI'26] NamedCurves: Learned Image Enhancement via Color Namingβ37May 26, 2026Updated 2 months ago
- [ECCV-W] Official repo for the paper "ComiCap: A VLMs pipeline for dense captioning of Comic Panels"β15Nov 20, 2024Updated last year
- Text-DIAE: A Self-Supervised Degradation Invariant Autoencoders for Text Recognition and Document Enhancement - AAAI 2023β30Jul 12, 2023Updated 3 years ago
- A Bottom-Up Instance Segmentation Strategy for segmenting document instances using Transformersβ59Sep 9, 2024Updated last year
- [ICCV 2023] - Zero-shot Composed Image Retrieval with Textual Inversionβ198Jul 31, 2025Updated last year
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [ICIAP 2023] Learning Landmarks Motion from Speech for Speaker-Agnostic 3D Talking Heads Generationβ59Dec 12, 2023Updated 2 years ago
- [ICCV 2025] - Image Intrinsic Scale Assessment: Bridging the Gap Between Quality and Resolutionβ17Aug 16, 2025Updated last year
- ICDAR 2019β25Aug 2, 2019Updated 7 years ago
- Generate a transcript for your favourite Manga: Detect manga characters, text blocks and panels. Order panels. Cluster characters. Match β¦β467Jun 27, 2025Updated last year
- Based on the WACV 2020 paper - Fine Grained Classification and Retrieval by Combining Visual and Locally Pooled Textual Featuresβ25Nov 15, 2021Updated 4 years ago
- Quality-Aware Image-Text Alignment for Opinion-Unaware Image Quality Assessmentβ132Mar 10, 2025Updated last year
- OCR Annotations from Amazon Textract for Industry Documents Libraryβ104Aug 20, 2022Updated 3 years ago
- Implementation on pytorch of the code from the ECCV 2018 paper - Single Shot Scene Text Retrievalβ13Dec 15, 2021Updated 4 years ago
- TextAdaIN: Paying Attention to Shortcut Learning in Text Recognizersβ21Jul 26, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Official repository of the paper: "A Comprehensive Gold Standard and Benchmark for Comics Text Detection and Recognition"β27Jul 10, 2023Updated 3 years ago
- [ICLR 2026] - Spectral Concept Selection and Cross-modal Representation Learning for Generalized Category Discoveryβ23Mar 18, 2026Updated 5 months ago
- Code for "A Comprehensive Empirical Evaluation on Online Continual Learning" ICCVW 2023 VCL Workshopβ46Apr 8, 2024Updated 2 years ago
- DocEnTr: An end-to-end document image enhancement transformer - ICPR 2022β190Jan 17, 2025Updated last year
- Official repository for our paper on "Attribution-aware Weight Transfer: A Warm-Start Initialization for Class-Incremental Semantic Segmeβ¦β12Jan 3, 2023Updated 3 years ago
- Multimodal Agentic Document QA benchmark (MADQA)β41Mar 13, 2026Updated 5 months ago
- [WACV 2024 Oral] - ARNIQA: Learning Distortion Manifold for Image Quality Assessmentβ156Jun 18, 2026Updated 2 months ago