Enhancing Image Classification Accuracy Based on AlexNet
DOI:
https://doi.org/10.54097/9c8kkn59Keywords:
AlexNet Model, Neural Network, Picture Classification, CIFAR-10 Dataset.Abstract
This study explores the potential of deep learning for small-scale image classification tasks through the utilization of the classical AlexNet model on the CIFAR-10 dataset. The methodology involves meticulous architectural adjustments, comprehensive data preprocessing, and strategic optimization of the learning rate decay, resulting in substantial improvements in model performance. In addition, innovative data augmentation techniques are introduced to enhance the model's robustness and generalization capabilities. The experimental outcomes unequivocally underscore the efficacy of deep learning in addressing small-scale image classification challenges, offering versatile applications across diverse domains such as image recognition, autonomous driving, and medical diagnostics. Acknowledging the dynamic nature of deep learning, ongoing research endeavors are warranted. Future directions may encompass the exploration of more intricate neural network architectures, advanced data augmentation methodologies, and the implementation of interpretability tools to augment both performance and comprehensibility. This study provides compelling evidence of deep learning's aptitude for small-scale image classification tasks, offering valuable insights and guidance for future research and practical implementations. As the field of deep learning continues to evolve, this work aims to serve as a valuable reference and catalyst for further advancements in the domain.
Downloads
References
Krizhevsky A, Sutskever I, Hinton G E. ImageNet classification with deep convolutional neural networks [J]. Communications of the ACM, 2017, 60 (6): 84 - 90.
Simonyan K, Zisserman A. Very deep convolutional networks for large-scale image recognition [J]. arXiv preprint arXiv: 1409. 1556, 2014.
Szegedy C, Liu W, Jia Y, et al. Going deeper with convolutions [C]//Proceedings of the IEEE conference on computer vision and pattern recognition. 2015: 1 - 9.
He K, Zhang X, Ren S, et al. Deep residual learning for image recognition [C]//Proceedings of the IEEE conference on computer vision and pattern recognition. 2016: 770 - 778.
Szegedy C, Ioffe S, Vanhoucke V, et al. Inception-v4, inception-resnet and the impact of residual connections on learning [C]//Proceedings of the AAAI conference on artificial intelligence. 2017, 31 (1).
Krizhevsky A, Hinton G. Learning multiple layers of features from tiny images [J]. 2009.
Krizhevsky A, Nair V, Hinton G. Cifar-10 (canadian institute for advanced research) [J]. URL http://www. cs. toronto. edu/kriz/cifar. html, 2010, 5 (4): 1.
Szegedy C, Liu W, Jia Y, et al. Going deeper with convolutions[C]//Proceedings of the IEEE conference on computer vision and pattern recognition. 2015: 1 - 9.
Zagoruyko S, Komodakis N. Wide residual networks [J]. arXiv preprint arXiv: 1605.07146, 2016.
Taigman Y, Yang M, Ranzato M A, et al. Deepface: Closing the gap to human-level performance in face verification [C]//Proceedings of the IEEE conference on computer vision and pattern recognition. 2014: 1701 - 1708.
Simonyan K, Zisserman A. Very deep convolutional networks for large-scale image recognition [J]. arXiv preprint arXiv: 1409. 1556, 2014.
Downloads
Published
Issue
Section
License
Copyright (c) 2024 Highlights in Science, Engineering and Technology

This work is licensed under a Creative Commons Attribution-NonCommercial 4.0 International License.







