ImageNet classification with deep convolutional neural networks
Классификация ImageNet с использованием глубоких сверточных нейронных сетей
2017-05-24
SCID: 54.1/xbbm6eum
Discuss with AI
GPU-accelerated convolutionImageNet classificationdeep convolutional neural networksdropout regularizationtop-5 error rate
Figures from the paper
Abstract (AI)
We trained a large, deep convolutional neural network to classify the 1.2 million high-resolution images in the ImageNet LSVRC-2010 contest into the 1000 different classes. On the test data, we achieved top-1 and top-5 error rates of 37.5% and 17.0%, respectively, which is considerably better than the previous state-of-the-art. The neural network, which has 60 million parameters and 650,000 neurons, consists of five convolutional layers, some of which are followed by max-pooling layers, and three fully connected layers with a final 1000-way softmax. To make training faster, we used non-saturating neurons and a very efficient GPU implementation of the convolution operation. To reduce overfitting in the fully connected layers we employed a recently developed regularization method called "dropout" that proved to be very effective. We also entered a variant of this model in the ILSVRC-2012 competition and achieved a winning top-5 test error rate of 15.3%, compared to 26.2% achieved by the second-best entry.
Key Findings
1
A deep convolutional neural network classified 1.2 million ImageNet images across 1,000 classes, achieving 37.5% top-1 and 17.0% top-5 error rates.
2
A model variant won ILSVRC-2012 with a 15.3% top-5 error rate, outperforming the second-best entry’s 26.2%.
3
Dropout regularization effectively reduced overfitting in the fully connected layers.
4
Non-saturating neurons and an efficient GPU convolution implementation substantially accelerated network training.
5
The 60-million-parameter architecture used five convolutional layers, max-pooling, three fully connected layers, and a final 1,000-way softmax.
Research Object
A large, deep convolutional neural network trained on the ImageNet (ILSVRC) dataset
Research Subject
Deep convolutional neural network classification performance, including top-1 and top-5 error rates and the effects of network architecture and training techniques
Publication Details
Publication Date
2017-05-24
Journal
Publisher
ISSN
Open access PDF
Access Type
Author Information
Download PDF
Subscribe to digest
References available in scid.ai4
Cited by20
Swin Transformer: Hierarchical Vision Transformer using Shifted Windows2021
A Survey of Convolutional Neural Networks: Analysis, Applications, and Prospects2021
Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without Convolutions2021
Video Swin Transformer2022
Interpreting Black-Box Models: A Review on Explainable Artificial Intelligence2023
Hyperspectral Image Classification—Traditional to Deep Models: A Survey for Future Prospects2021
Artificial Neural Network Algorithms for 3D Printing2020
An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale2020
A review of vibration-based damage detection in civil structures: From traditional methods to Machine Learning and Deep Learning applications2020
ECA-Net: Efficient Channel Attention for Deep Convolutional Neural Networks2020
Albumentations: Fast and Flexible Image Augmentations2020
Introduction to Radiomics2020
Graph neural networks: A review of methods and applications2020
A Survey of Autonomous Driving: Common Practices and Emerging Technologies2020
Digital Twin: Values, Challenges and Enablers From a Modeling Perspective2020
A survey on semi-supervised learning2019
Graph convolutional networks: a comprehensive review2019
Deep Learning for Generic Object Detection: A Survey2019
Voxceleb: Large-scale speaker verification in the wild2019
Dynamic Graph CNN for Learning on Point Clouds2019