Geometry-Aware Scene Text Detection with Instance Transformation Network

Fangfang Wang, Liming Zhao, Xi Li, Xinchao Wang, Dacheng Tao
2018 2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition  
Localizing text in the wild is challenging in the situations of complicated geometric layout of the targets like random orientation and large aspect ratio. In this paper, we propose a geometry-aware modeling approach tailored for scene text representation with an end-to-end learning scheme. In our approach, a novel Instance Transformation Network (ITN) is presented to learn the geometry-aware representation encoding the unique geometric configurations of scene text instances with in-network
more » ... sformation embedding, resulting in a robust and elegant framework to detect words or text lines at one pass. An end-to-end multi-task learning strategy with transformation regression, text/non-text classification and coordinates regression is adopted in the ITN. Experiments on the benchmark datasets demonstrate the effectiveness of the proposed approach in detecting scene text in various geometric configurations. * Authors contributed equally,
doi:10.1109/cvpr.2018.00150 dblp:conf/cvpr/WangZ0WT18 fatcat:z2ys5ri4ijhgdea3yrpaxpkode