AAAI Publications, Thirty-First AAAI Conference on Artificial Intelligence

Font Size: 
TextBoxes: A Fast Text Detector with a Single Deep Neural Network
Minghui Liao, Baoguang Shi, Xiang Bai, Xinggang Wang, Wenyu Liu

Last modified: 2017-02-12


This paper presents an end-to-end trainable fast scene text detector, named TextBoxes, which detects scene text with both high accuracy and efficiency in a single network forward pass, involving no post-process except for a standard non-maximum suppression. TextBoxes outperforms competing methods in terms of text localization accuracy and is much faster, taking only 0.09s per image in a fast implementation. Furthermore, combined with a text recognizer, TextBoxes significantly outperforms state-of-the-art approaches on word spotting and end-to-end text recognition tasks.


scene text; convolutional neural network; text localization; word spotting; end to end recognition;

Full Text: PDF