next up previous contents
Next: Intriguing properties of neural Up: Summary of References Related Previous: 3D convolutional neural networks   Contents

Subsections

Multi-digit Number Recognition from Street View Imagery using Deep Convolutional Neural Networks [28]

Original Abstract

Recognizing arbitrary multi-character text in unconstrained natural photographs is a hard problem. In this paper, we address an equally hard sub-problem in this domain viz. recognizing arbitrary multi-digit numbers from Street View imagery. Traditional approaches to solve this problem typically separate out the localization, segmentation, and recognition steps. In this paper we propose a unified approach that integrates these three steps via the use of a deep convolutional neural network that operates directly on the image pixels. We employ the DistBelief implementation of deep neural networks in order to train large, distributed neural networks on high quality images. We find that the performance of this approach increases with the depth of the convolutional network, with the best performance occurring in the deepest architecture we trained, with eleven hidden layers. We evaluate this approach on the publicly available SVHN dataset and achieve over 96


next up previous contents
Next: Intriguing properties of neural Up: Summary of References Related Previous: 3D convolutional neural networks   Contents
Miquel Perello Nieto 2014-11-28