The official repo of the Comics Survey: "A missing piece in Vision and Language: A Survey on Comics Understanding"
β142Jan 2, 2025Updated last year
Alternatives and similar repositories for awesome-comics-understanding
Users that are interested in awesome-comics-understanding are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- β32May 28, 2025Updated last year
- [ICDAR 2024] (Best Student Paperπ) Exploring Knowledge Distillation Towards Document Object Detection with Structured Graph Creationβ14Sep 6, 2024Updated 2 years ago
- Repository for "CoMix: Comprehensive Benchmark for Multi-Task Comic Understanding"β18Nov 20, 2024Updated last year
- [ECCV 2024] - Improving Zero-shot Generalization of Learned Prompts via Unsupervised Knowledge Distillationβ62Feb 20, 2026Updated 6 months ago
- [CVPR 2024 Highlight] - Stationary Representations: Optimally Approximating Compatibility and Implications for Improved Model Replacementβ¦β14Oct 21, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Comics Dataset Framework for Comics Understandingβ43Sep 1, 2025Updated last year
- [WACV 2026 Round 1] Beyond Single Object Text-to-SVG Synthesis with Comprehensive Canvas Layoutβ22Oct 11, 2025Updated 10 months ago
- Hadwritten Text Recognition in Few-shot Scenarioβ22Mar 25, 2023Updated 3 years ago
- This repo contains the code of "Contrastive Supervised Distillation for Continual Representation Learning", Tommaso Barletti, NiccolΓ² Bioβ¦β20Jul 5, 2022Updated 4 years ago
- WACV 2022 Paper - Is An Image Worth Five Sentences? A New Look into Semantics for Image-Text Matchingβ16Dec 10, 2021Updated 4 years ago
- Doc2Graph transforms documents into graphs and exploit a GNN to solve several tasks.β140Oct 18, 2025Updated 10 months ago
- [WACV 2024] - Reference-based Restoration of Digitized Analog Videotapesβ61Feb 11, 2024Updated 2 years ago
- [ICCVW 2023] - Mapping Memes to Words for Multimodal Hateful Meme Classificationβ28Apr 17, 2025Updated last year
- [CVPR 2023] NEFER a Dataset for Neuromorphic Event-based Facial Expression Recognitionβ30Dec 7, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICCV 2023] - Composed Image Retrieval on Common Objects in context (CIRCO) datasetβ87Aug 6, 2025Updated last year
- Optocal Character Recognition (OCR / HTR) using Transformersβ11Aug 20, 2022Updated 4 years ago
- [IEEE TMM 2023] This is the official repo of the paper "Perceptual Quality Improvement in Videoconferencing using Keyframes-based GAN".β17Dec 10, 2024Updated last year
- β28Mar 7, 2025Updated last year
- Official evaluation scripts and baseline prompts for the DocVQA 2026 (ICDAR 2026) Competition on Multimodal Reasoning over Documents.β18Mar 16, 2026Updated 5 months ago
- Official PyTorch Implementation of DocSynth: A Layout Guided Approach for Controllable Document Image Synthesis - ICDAR 2021β95Jul 16, 2021Updated 5 years ago
- Official Pytorch code for MANTRA - Memory Augmented Neural Trajectory Predictor (CVPR2020)β77Aug 24, 2022Updated 4 years ago
- [ECCV'24] [TPAMI'26] NamedCurves: Learned Image Enhancement via Color Namingβ38May 26, 2026Updated 3 months ago
- [ECCV-W] Official repo for the paper "ComiCap: A VLMs pipeline for dense captioning of Comic Panels"β15Nov 20, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Text-DIAE: A Self-Supervised Degradation Invariant Autoencoders for Text Recognition and Document Enhancement - AAAI 2023β30Jul 12, 2023Updated 3 years ago
- A Bottom-Up Instance Segmentation Strategy for segmenting document instances using Transformersβ59Sep 9, 2024Updated last year
- [ICIAP 2023] Learning Landmarks Motion from Speech for Speaker-Agnostic 3D Talking Heads Generationβ59Dec 12, 2023Updated 2 years ago
- β18Sep 14, 2024Updated last year
- [ICCV 2025] - Image Intrinsic Scale Assessment: Bridging the Gap Between Quality and Resolutionβ17Aug 16, 2025Updated last year
- ICDAR 2019β25Aug 2, 2019Updated 7 years ago
- Generate a transcript for your favourite Manga: Detect manga characters, text blocks and panels. Order panels. Cluster characters. Match β¦β471Jun 27, 2025Updated last year
- Attempt at reproducing the metric from Neurips 2023 Unlearning Challenge on Kaggle. Code for training checkpoints on retain set and unleaβ¦β12Nov 8, 2023Updated 2 years ago
- Quality-Aware Image-Text Alignment for Opinion-Unaware Image Quality Assessmentβ132Mar 10, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- OCR Annotations from Amazon Textract for Industry Documents Libraryβ105Aug 20, 2022Updated 4 years ago
- Implementation on pytorch of the code from the ECCV 2018 paper - Single Shot Scene Text Retrievalβ13Dec 15, 2021Updated 4 years ago
- TextAdaIN: Paying Attention to Shortcut Learning in Text Recognizersβ21Jul 26, 2022Updated 4 years ago
- Official repository of the paper: "A Comprehensive Gold Standard and Benchmark for Comics Text Detection and Recognition"β27Jul 10, 2023Updated 3 years ago
- [ICLR 2026] - Spectral Concept Selection and Cross-modal Representation Learning for Generalized Category Discoveryβ23Mar 18, 2026Updated 5 months ago
- Code for "A Comprehensive Empirical Evaluation on Online Continual Learning" ICCVW 2023 VCL Workshopβ46Apr 8, 2024Updated 2 years ago
- DocEnTr: An end-to-end document image enhancement transformer - ICPR 2022β191Jan 17, 2025Updated last year