paper-with-me

Papers

Advancing Manga Analysis: Comprehensive Segmentation Annotations for the Manga109 Dataset

2025-01-01 · CVPR 2025 1 · Minshan Xie, Jian Lin, Hanyuan Liu, Chengze Li, Tien-Tsin Wong

Manga, a popular form of multimodal artwork, has traditionally been overlooked in deep learning advancements due to the absence of a robust dataset and comprehensive annotation. Manga segmentation is the key to the digital migration of manga. There exists a significant domain gap between the manga and the natural images, that fails most existing learning-based methods. To address this gap, we introduce an augmented segmentation annotation for the Manga109 dataset, a collection of 109 manga volumes, that offers intricate artworks in a rich variety of styles. We introduce a detailed annotation that extends beyond the original simple bounding boxes to the segmentation masks with pixel-level precision. It provides object category, location, and instance information that can be used for semantic segmentation and instance segmentation. We also provide a comprehensive analysis of our annotation dataset from various aspects. We further measure the improvement of the state-of-the-art segmentation model after training it with our augmented dataset. The benefits of this augmented dataset are profound, with the potential to significantly enhance manga analysis algorithms and catalyze the novel development in digital art processing and cultural analytics. This annotation, named MangaSeg, is publicly available at https://huggingface.co/datasets/MS92/MangaSegmentation.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Instance SegmentationSegmentationSemantic Segmentation

Similar Papers 제목 키워드 기반

MangaVQA and MangaLMM: A Benchmark and Specialized Model for Multimodal Manga Understanding

2025-05-26 · Jeonghun Baek, Kazuki Egashira, Shota Onohara, Atsuyuki Miyai 외

Manga, or Japanese comics, is a richly multimodal narrative form that blends images and text in complex ways. Teaching large multimodal models (LMMs) to understand such narratives at a human-like level could help manga c…

Question AnsweringVisual Question Answering

Manga109-v2026: Revisiting Manga109 Annotations for Modern Manga Understanding

2026-05-20 · Jeonghun Baek, Atsuyuki Miyai, Shota Onohara, Hikaru Ikuta 외 arxiv

Manga is a culturally distinctive multimodal medium and one of the most influential forms of Japanese popular culture. As AI systems increasingly target manga understanding, OCR, and translation, Manga109 has become a fo…

Building a Manga Dataset "Manga109" with Annotations for Multimedia Applications

2020-05-09 · Kiyoharu Aizawa, Azuma Fujimoto, Atsushi Otsubo, Toru Ogawa 외

Manga, or comics, which are a type of multimodal artwork, have been left behind in the recent trend of deep learning applications because of the lack of a proper dataset. Hence, we built Manga109, a dataset consisting of…

Retrieval

COT-AD: Cotton Analysis Dataset

2025-07-24 · Akbar Ali, Mahek Vyas, Soumyaratna Debnath, Chanda Grover Kamra 외 arxiv

This paper presents COT-AD, a comprehensive Dataset designed to enhance cotton crop analysis through computer vision. Comprising over 25,000 images captured throughout the cotton growth cycle, with 5,000 annotated images…

Image Restoration

Unconstrained Text Detection in Manga: a New Dataset and Baseline

2020-09-09 · Julián Del Gobbo, Rosana Matuk Herrera

The detection and recognition of unconstrained text is an open problem in research. Text in comic books has unusual styles that raise many challenges for text detection. This work aims to binarize text in a comic genre w…

BinarizationText Detection