NE CVOct 30, 2017

Evolving Deep Convolutional Neural Networks for Image Classification

Yanan Sun, Bing Xue, Mengjie Zhang, Gary G. Yen

arXiv:1710.10741v329.6651 citations

Originality Incremental advance

AI Analysis

This addresses the problem of inefficient evolution for deep networks in image classification, offering a novel approach that is incremental but shows strong performance gains.

The paper tackles the challenge of scaling evolutionary computation to deep neural networks by proposing a genetic algorithm to evolve architectures and weight initializations for image classification, achieving superior classification error rates and parameter efficiency compared to 22 existing state-of-the-art methods.

Evolutionary computation methods have been successfully applied to neural networks since two decades ago, while those methods cannot scale well to the modern deep neural networks due to the complicated architectures and large quantities of connection weights. In this paper, we propose a new method using genetic algorithms for evolving the architectures and connection weight initialization values of a deep convolutional neural network to address image classification problems. In the proposed algorithm, an efficient variable-length gene encoding strategy is designed to represent the different building blocks and the unpredictable optimal depth in convolutional neural networks. In addition, a new representation scheme is developed for effectively initializing connection weights of deep convolutional neural networks, which is expected to avoid networks getting stuck into local minima which is typically a major issue in the backward gradient-based optimization. Furthermore, a novel fitness evaluation method is proposed to speed up the heuristic search with substantially less computational resource. The proposed algorithm is examined and compared with 22 existing algorithms on nine widely used image classification tasks, including the state-of-the-art methods. The experimental results demonstrate the remarkable superiority of the proposed algorithm over the state-of-the-art algorithms in terms of classification error rate and the number of parameters (weights).

View on arXiv PDF

Similar