TASK	DATASET	MODEL	METRIC NAME	METRIC VALUE	GLOBAL RANK	REMOVE
Text-Line Extraction	DIVA-HisDB	Semantic Seg Preprocessing	Line IoU	99.42	# 1
Text-Line Extraction	DIVA-HisDB	Semantic Seg Preprocessing	Pixel IoU	96.11	# 1

Badge	Markdown
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/labeling-cutting-grouping-an-efficient-text/text-line-extraction-on-diva-hisdb)](https://paperswithcode.com/sota/text-line-extraction-on-diva-hisdb?p=labeling-cutting-grouping-an-efficient-text)`

Labeling, Cutting, Grouping: an Efficient Text Line Segmentation Method for Medieval Manuscripts

11 Jun 2019 · Michele Alberti, Lars Vögtlin, Vinaychandran Pondenkandath, Mathias Seuret, Rolf Ingold, Marcus Liwicki ·

This paper introduces a new way for text-line extraction by integrating deep-learning based pre-classification and state-of-the-art segmentation methods. Text-line extraction in complex handwritten documents poses a significant challenge, even to the most modern computer vision algorithms. Historical manuscripts are a particularly hard class of documents as they present several forms of noise, such as degradation, bleed-through, interlinear glosses, and elaborated scripts. In this work, we propose a novel method which uses semantic segmentation at pixel level as intermediate task, followed by a text-line extraction step. We measured the performance of our method on a recent dataset of challenging medieval manuscripts and surpassed state-of-the-art results by reducing the error by 80.7%. Furthermore, we demonstrate the effectiveness of our approach on various other datasets written in different scripts. Hence, our contribution is two-fold. First, we demonstrate that semantic pixel segmentation can be used as strong denoising pre-processing step before performing text line extraction. Second, we introduce a novel, simple and robust algorithm that leverages the high-quality semantic segmentation to achieve a text-line extraction performance of 99.42% line IU on a challenging dataset.

PDF Abstract

Code

Add Remove Mark official

DIVA-DIA/Text-Line-Segmentation-Met… official

Tasks

Add Remove

Denoising

Segmentation

Semantic Segmentation

Text-Line Extraction

Datasets

DIVA-HisDB

Results from the Paper

Edit

Ranked #1 on Text-Line Extraction on DIVA-HisDB

Get a GitHub badge

Task	Dataset	Model	Metric Name	Metric Value	Global Rank	Result	Benchmark
Text-Line Extraction	DIVA-HisDB	Semantic Seg Preprocessing	Line IoU	99.42	# 1		Compare
Text-Line Extraction	DIVA-HisDB	Semantic Seg Preprocessing	Pixel IoU	96.11	# 1		Compare

Methods

Add Remove

No methods listed for this paper. Add relevant methods here

Edit Social Preview

Labeling, Cutting, Grouping: an Efficient Text Line Segmentation Method for Medieval Manuscripts

Code Edit Add Remove Mark official

Tasks Edit Add Remove

Datasets Edit

Results from the Paper Edit

Methods Edit Add Remove

Code

Add Remove Mark official

Tasks

Add Remove

Datasets

Results from the Paper

Edit

Methods

Add Remove