The official repo of the Comics Survey: "A missing piece in Vision and Language: A Survey on Comics Understanding"
β141Jan 2, 2025Updated last year
Alternatives and similar repositories for awesome-comics-understanding
Users that are interested in awesome-comics-understanding are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- β32May 28, 2025Updated last year
- [ICDAR 2024] (Best Student Paperπ) Exploring Knowledge Distillation Towards Document Object Detection with Structured Graph Creationβ15Sep 6, 2024Updated 2 years ago
- Repository for "CoMix: Comprehensive Benchmark for Multi-Task Comic Understanding"β18Nov 20, 2024Updated last year
- [ECCV 2024] - Improving Zero-shot Generalization of Learned Prompts via Unsupervised Knowledge Distillationβ62Feb 20, 2026Updated 7 months ago
- [CVPR 2024 Highlight] - Stationary Representations: Optimally Approximating Compatibility and Implications for Improved Model Replacementβ¦β14Oct 21, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [WACV 2026 Round 1] Beyond Single Object Text-to-SVG Synthesis with Comprehensive Canvas Layoutβ24Oct 11, 2025Updated 11 months ago
- [WACV 2024] - Reference-based Restoration of Digitized Analog Videotapesβ62Feb 11, 2024Updated 2 years ago
- [ICCVW 2023] - Mapping Memes to Words for Multimodal Hateful Meme Classificationβ28Apr 17, 2025Updated last year
- [CVPR 2023] NEFER a Dataset for Neuromorphic Event-based Facial Expression Recognitionβ30Dec 7, 2023Updated 2 years ago
- [ICCV 2023] - Composed Image Retrieval on Common Objects in context (CIRCO) datasetβ87Aug 6, 2025Updated last year
- Optocal Character Recognition (OCR / HTR) using Transformersβ11Aug 20, 2022Updated 4 years ago
- [IEEE TMM 2023] This is the official repo of the paper "Perceptual Quality Improvement in Videoconferencing using Keyframes-based GAN".β17Dec 10, 2024Updated last year
- β28Mar 7, 2025Updated last year
- Official PyTorch Implementation of DocSynth: A Layout Guided Approach for Controllable Document Image Synthesis - ICDAR 2021β95Jul 16, 2021Updated 5 years ago
- End-to-end encrypted email - Proton Mail β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Official Pytorch code for MANTRA - Memory Augmented Neural Trajectory Predictor (CVPR2020)β77Aug 24, 2022Updated 4 years ago
- [ECCV'24] [TPAMI'26] NamedCurves: Learned Image Enhancement via Color Namingβ38May 26, 2026Updated 4 months ago
- [ECCV-W] Official repo for the paper "ComiCap: A VLMs pipeline for dense captioning of Comic Panels"β15Nov 20, 2024Updated last year
- A Bottom-Up Instance Segmentation Strategy for segmenting document instances using Transformersβ59Sep 9, 2024Updated 2 years ago
- [ICCV 2023] - Zero-shot Composed Image Retrieval with Textual Inversionβ197Jul 31, 2025Updated last year
- [ICIAP 2023] Learning Landmarks Motion from Speech for Speaker-Agnostic 3D Talking Heads Generationβ59Dec 12, 2023Updated 2 years ago
- [ICCV 2025] - Image Intrinsic Scale Assessment: Bridging the Gap Between Quality and Resolutionβ18Aug 16, 2025Updated last year
- Generate a transcript for your favourite Manga: Detect manga characters, text blocks and panels. Order panels. Cluster characters. Match β¦β470Jun 27, 2025Updated last year
- Quality-Aware Image-Text Alignment for Opinion-Unaware Image Quality Assessmentβ132Mar 10, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI β’ AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- OCR Annotations from Amazon Textract for Industry Documents Libraryβ105Aug 20, 2022Updated 4 years ago
- Implementation on pytorch of the code from the ECCV 2018 paper - Single Shot Scene Text Retrievalβ13Dec 15, 2021Updated 4 years ago
- TextAdaIN: Paying Attention to Shortcut Learning in Text Recognizersβ21Jul 26, 2022Updated 4 years ago
- Official repository of the paper: "A Comprehensive Gold Standard and Benchmark for Comics Text Detection and Recognition"β27Jul 10, 2023Updated 3 years ago
- [ICLR 2026] - Spectral Concept Selection and Cross-modal Representation Learning for Generalized Category Discoveryβ23Mar 18, 2026Updated 6 months ago
- Code for "A Comprehensive Empirical Evaluation on Online Continual Learning" ICCVW 2023 VCL Workshopβ46Apr 8, 2024Updated 2 years ago
- Multimodal Agentic Document QA benchmark (MADQA)β42Mar 13, 2026Updated 6 months ago
- [WACV 2024 Oral] - ARNIQA: Learning Distortion Manifold for Image Quality Assessmentβ157Jun 18, 2026Updated 3 months ago
- Official PyTorch Implementation of "Rethinking HTG Evaluation: Bridging Generation and Recognition" (Oral) - 1st Workshop on Critical Evaβ¦β17Sep 23, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI β’ AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- The Land-Diffuser is a novel application of the Denoising Diffusion Probabilistic Model (DDPM) in the realm of 3D Talking Head generationβ¦β13Dec 23, 2023Updated 2 years ago
- Scene Text Aware Cross Modal Retrieval (StacMR)β24Sep 3, 2021Updated 5 years ago
- Papers, datasets, and resources related to 2D cartoon video research. Contributions welcome.β217Aug 25, 2026Updated last month
- β17Jul 11, 2024Updated 2 years ago
- Searching a High Performance Feature Extractor for Text Recognition Network. TPAMI 2022β13Nov 25, 2022Updated 3 years ago
- Official Repository of RefChartQA: Grounding Visual Answer on Chart Images through Instruction Tuningβ15Jul 9, 2025Updated last year
- RoDLA: Benchmarking the Robustness of Document Layout Analysis Modelsβ40Mar 26, 2025Updated last year