Page segmentation for historical document images based on superpixel classification with unsupervised feature learning

Chen, Kai; Liu, Cheng-Lin; Seuret, Mathias; Liwicki, Marcus; Hennebert, Jean; Ingold, Rolf

doi:10.1109/DAS.2016.13

Chen, Kai; Liu, Cheng-Lin; Seuret, Mathias; Liwicki, Marcus; Hennebert, Jean; Ingold, Rolf

2016

Télécharger

Formats

Format
BibTeX
MARCXML
TextMARC
MARC
DublinCore
EndNote
NLM
RefWorks
RIS

Résumé

In this paper, we present an efficient page segmentation method for historical document images. Many existing methods either rely on hand-crafted features or perform rather slow as they treat the problem as a pixel-level assignment problem. In order to create a feasible method for real applications, we propose to use superpixels as basic units of segmentation, and features are learned directly from pixels. An image is first oversegmented into superpixels with the simple linear iterative clustering (SLIC) algorithm. Then, each superpixel is represented by the features of its central pixel. The features are learned from pixel intensity values with stacked convolutional autoencoders in an unsupervised manner. A support vector machine (SVM) classifier is used to classify superpixels into four classes: periphery, background, text block, and decoration. Finally, the segmentation results are refined by a connected component based smoothing procedure. Experiments on three public datasets demonstrate that compared to our previous method, the proposed method is much faster and achieves comparable segmentation results. Additionally, much fewer pixels are used for classifier training.

Détails

Titre Page segmentation for historical document images based on superpixel classification with unsupervised feature learning

Auteur(s)/ trice(s) Chen, Kai (University of Fribourg, Fribourg, Switzerland)
Liu, Cheng-Lin (Institute of Automation of Chinese Academy of Sciences, China)
Seuret, Mathias (University of Fribourg, Fribourg, Switzerland)
Liwicki, Marcus (University of Fribourg, Fribourg, Switzerland)
Hennebert, Jean (University of Fribourg, Fribourg, Switzerland ; School of Engineering and Architecture (HEIA-FR), HES-SO University of Applied Sciences Western Switzerland)
Ingold, Rolf (University of Fribourg, Fribourg, Switzerland)

Date 2016-04

Publié dans Proceedings of the 2016 12th IAPR Workshop on Document Analysis Systems (DAS), 11-14 April 2016, Santorini, Greece

Volume pp. 299-304

Editeur Santorini, Greece, 11-14 April 2016

Pagination 6 p.

Présenté à 2016 12th IAPR Workshop on Document Analysis Systems (DAS), Santorini, Greece, 2016-04-11, 2016-04-16

ISBN 978-1-5090-1792-8

DOI https://doi.org/10.1109/DAS.2016.13

Mots-clés (libres) image segmentation ; training ; feature extraction ; support vector machines ; clustering algorithms ; classification algorithms ; labeling ; page segmentation ; layout analysis ; hisotrical document image ; superpixel ; SLIC ; autoencoder

Type de papier published full paper

Domaine Design et Arts visuels
Ingénierie et Architecture

Ecole HEIA-FR

Institut iCoSys - Institut des systèmes complexes

Le document apparaît dans Documents de conférences
Global

Résumé

Détails

Actions