Manga109Dialog A Large-scale Dialogue Dataset for Comics Speaker Detection

06/30/2023
by   Yingxuan Li, et al.
0

The expanding market for e-comics has spurred interest in the development of automated methods to analyze comics. For further understanding of comics, an automated approach is needed to link text in comics to characters speaking the words. Comics speaker detection research has practical applications, such as automatic character assignment for audiobooks, automatic translation according to characters' personalities, and inference of character relationships and stories. To deal with the problem of insufficient speaker-to-text annotations, we created a new annotation dataset Manga109Dialog based on Manga109. Manga109Dialog is the world's largest comics speaker annotation dataset, containing 132,692 speaker-to-text pairs. We further divided our dataset into different levels by prediction difficulties to evaluate speaker detection methods more appropriately. Unlike existing methods mainly based on distances, we propose a deep learning-based method using scene graph generation models. Due to the unique features of comics, we enhance the performance of our proposed model by considering the frame reading order. We conducted experiments using Manga109Dialog and other datasets. Experimental results demonstrate that our scene-graph-based approach outperforms existing methods, achieving a prediction accuracy of over 75

READ FULL TEXT

page 2

page 4

page 7

page 9

research
09/10/2022

IR-LPR: Large Scale of Iranian License Plate Recognition Dataset

Object detection has always been practical. There are so many things in ...
research
09/18/2022

A Benchmark for Understanding and Generating Dialogue between Characters in Stories

Many classical fairy tales, fiction, and screenplays leverage dialogue t...
research
04/03/2019

Character Region Awareness for Text Detection

Scene text detection methods based on neural networks have emerged recen...
research
06/07/2022

The Influence of Dataset Partitioning on Dysfluency Detection Systems

This paper empirically investigates the influence of different data spli...
research
01/02/2019

Detecting Text in the Wild with Deep Character Embedding Network

Most text detection methods hypothesize texts are horizontal or multi-or...
research
06/09/2022

TwiBot-22: Towards Graph-Based Twitter Bot Detection

Twitter bot detection has become an increasingly important task to comba...
research
09/25/2013

Characterness: An Indicator of Text in the Wild

Text in an image provides vital information for interpreting its content...

Please sign up or login with your details

Forgot password? Click here to reset