DACyTAr - Datos Primarios en Acceso Abierto de la Ciencia y la Tecnología Argentina

Unconstrained Text Detection in Manga: a New Dataset and Baseline [Data set]

Compartir en
redes sociales

Registro completo

Título: Unconstrained Text Detection in Manga: a New Dataset and Baseline [Data set]

Autor(es): Matuk Herrera, Rosana; Del Gobbo, Julián

Afiliación(es) del/de los autor(es): Matuk Herrera, Rosana. Departamento de Ciencias Básicas. Universidad Nacional de Luján; Argentina

Resumen: "Unconstrained Text Detection in Manga: a New Dataset and Baseline". It contains 450 images with the text segmentation of images from Manga109 dataset (need to request access to this dataset in order to view original manga image). Pre-processed version of the images is how they were saved straight out of GIMP. These were later processed before using for training. Post-processed version of the images is after automatically removing small connected components and filling small holes. They are also slightly bigger in width/height in order to be multiples of 8. The text is split in 2 colors: black and pink. Text in black represents text we consider easy to recognize, which is mostly when inside a speech bubble. Text in pink represents text we consider harder to detect, such as text in covers, sound effects or text outside speech bubbles. Further details can be found in our paper.

Año de publicación: 2021

Idioma: inglés
indeterminado

Formato (Tipo MIME): application/zip
application/octet-stream

Clasificación temática de acuerdo a la FORD: Ciencias informáticas y de la información

Materia: Manga; Dataset; Segmentation;

Condiciones de uso: Disponible en acceso abierto bajo licencia Creative Commons https://creativecommons.org/licenses/by/2.5/ar/

Repositorio digital: REDIUNLU (UNLu) - Universidad Nacional de Luján

Acceder

Citación

Matuk Herrera, Rosana Del Gobbo, Julián (2021): Unconstrained Text Detection in Manga: a New Dataset and Baseline [Data set]. Universidad Nacional de Luján, https://doi.org/10.5281/zenodo.4511796.

Exportar cita

Archivo RDF/XML

Archivo JSON