EMNLP 20250 citations

ComicScene154: A Scene Dataset for Comic Analysis

Sandro Paval, Pascal Mei{\ss}ner, Ivan P. Yamshchikov

Abstract

Comics offer a compelling yet under-explored domain for computational narrative analysis, combining text and imagery in ways distinct from purely textual or audiovisual media. We introduce ComicScene154, a manually annotated dataset of scene-level narrative arcs derived from public-domain comic books spanning diverse genres. By conceptualizing comics as an abstraction for narrative-driven, multimodal data, we highlight their potential to inform broader research on multi-modal storytelling. To demonstrate the utility of ComicScene154, we present a baseline scene segmentation pipeline, providing an initial benchmark that future studies can build upon. Our results indicate that ComicScene154 constitutes a valuable resource for advancing computational methods in multimodal narrative understanding and expanding the scope of comic analysis within the Natural Language Processing community.

BibTeX
@inproceedings{emnlp2025_comicscene154asc,
  title = {ComicScene154: A Scene Dataset for Comic Analysis},
  author = {Sandro Paval and Pascal Mei{\ss}ner and Ivan P. Yamshchikov},
  booktitle = {EMNLP 2025},
  year = {2025}
}
ComicScene154: A Scene Dataset for Comic Analysis · EMNLP 2025