TASK	DATASET	MODEL	METRIC NAME	METRIC VALUE	GLOBAL RANK
Optical Character Recognition (OCR)	Benchmarking Chinese Text Recognition: Datasets, Baselines, and an Empirical Study	SRN	Accuracy (%)	65.0	# 4
Scene Text Recognition	ICDAR2013	SRN	Accuracy	95.5	# 21
Scene Text Recognition	SVT	SRN	Accuracy	91.5	# 20

Badge	Markdown
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/towards-accurate-scene-text-recognition-with/optical-character-recognition-on-benchmarking)](https://paperswithcode.com/sota/optical-character-recognition-on-benchmarking?p=towards-accurate-scene-text-recognition-with)`
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/towards-accurate-scene-text-recognition-with/scene-text-recognition-on-svt)](https://paperswithcode.com/sota/scene-text-recognition-on-svt?p=towards-accurate-scene-text-recognition-with)`
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/towards-accurate-scene-text-recognition-with/scene-text-recognition-on-icdar2013)](https://paperswithcode.com/sota/scene-text-recognition-on-icdar2013?p=towards-accurate-scene-text-recognition-with)`

Towards Accurate Scene Text Recognition with Semantic Reasoning Networks

CVPR 2020 · Deli Yu, Xuan Li, Chengquan Zhang, Junyu Han, Jingtuo Liu, Errui Ding ·

Scene text image contains two levels of contents: visual texture and semantic information. Although the previous scene text recognition methods have made great progress over the past few years, the research on mining semantic information to assist text recognition attracts less attention, only RNN-like structures are explored to implicitly model semantic information. However, we observe that RNN based methods have some obvious shortcomings, such as time-dependent decoding manner and one-way serial transmission of semantic context, which greatly limit the help of semantic information and the computation efficiency. To mitigate these limitations, we propose a novel end-to-end trainable framework named semantic reasoning network (SRN) for accurate scene text recognition, where a global semantic reasoning module (GSRM) is introduced to capture global semantic context through multi-way parallel transmission. The state-of-the-art results on 7 public benchmarks, including regular text, irregular text and non-Latin long text, verify the effectiveness and robustness of the proposed method. In addition, the speed of SRN has significant advantages over the RNN based methods, demonstrating its value in practical use.

PDF Abstract CVPR 2020 PDF CVPR 2020 Abstract

Code

Add Remove Mark official

PaddlePaddle/PaddleOCR

38,330

Media-Smart/vedastr

531

Tasks

Add Remove

Optical Character Recognition (OCR)

Scene Text Recognition

Datasets

ICDAR 2013

SVT

Results from the Paper

Edit

Ranked #4 on Optical Character Recognition (OCR) on Benchmarking Chinese Text Recognition: Datasets, Baselines, and an Empirical Study

Get a GitHub badge

Task	Dataset	Model	Metric Name	Metric Value	Global Rank	Benchmark
Optical Character Recognition (OCR)	Benchmarking Chinese Text Recognition: Datasets, Baselines, and an Empirical Study	SRN	Accuracy (%)	65.0	# 4	Compare
Scene Text Recognition	ICDAR2013	SRN	Accuracy	95.5	# 21	Compare
Scene Text Recognition	SVT	SRN	Accuracy	91.5	# 20	Compare

Methods

Add Remove

Semantic Reasoning Network • SPEED

Edit Social Preview

Towards Accurate Scene Text Recognition with Semantic Reasoning Networks

Code Edit Add Remove Mark official

Tasks Edit Add Remove

Datasets Edit

Results from the Paper Edit

Methods Edit Add Remove

Code

Add Remove Mark official

Tasks

Add Remove

Datasets

Results from the Paper

Edit

Methods

Add Remove