Hardware Accelerated Optimization of Deep Learning Model on Artificial Intelligence Chip
DOI:
https://doi.org/10.54097/fcis.v6i2.03Keywords:
Artificial Intelligence Chip, Hardware Accelerated, Deep LearningAbstract
With the rapid development of deep learning technology, the demand for computing resources is increasing, and the accelerated optimization of hardware on artificial intelligence (AI) chip has become one of the key ways to solve this challenge. This paper aims to explore the hardware acceleration optimization strategy of deep learning model on AI chip to improve the training and inference performance of the model. In this paper, the method and practice of optimizing deep learning model on AI chip are deeply analyzed by comprehensively considering the hardware characteristics such as parallel processing ability, energy-efficient computing, neural network accelerator, flexibility and programmability, high integration and heterogeneous computing structure. By designing and implementing an efficient convolution accelerator, the computational efficiency of the model is improved. The introduction of energy-efficient computing effectively reduces energy consumption, which provides feasibility for the practical application of mobile devices and embedded systems. At the same time, the optimization design of neural network accelerator becomes the core of hardware acceleration, and deep learning calculation such as convolution and matrix operation are accelerated through special hardware structure, which provides strong support for the real-time performance of the model. By analyzing the actual application cases of hardware accelerated optimization in different application scenarios, this paper highlights the key role of hardware accelerated optimization in improving the performance of deep learning model. Hardware accelerated optimization not only improves the computing efficiency, but also provides efficient and intelligent computing support for AI applications in different fields.
Downloads
References
Gupta, S. , Nguyen, D. , Rana, S. , Venkatesh, S. , & Kuttichira, D. P. (2022). Verification of integrity of deployed deep learning models using bayesian optimization. Knowledge-based systems (6), 241.
Khalifa, N. E. M., Taha, M. H. N. , Ali, D. E. , Slowik, A. , & Hassanien, A. E. (2020). Artificial intelligence technique for gene expression by tumor rna-seq data: a novel optimized deep learning approach. IEEE Access(8), 8.
Singh, K. , & Kapania, R. K. (2021). Accelerated optimization of curvilinearly stiffened panels using deep learning. Thin-Walled Structures, 161(3), 107418.
Geng, X. , Zhao, L. , Shi, L. , Yang, J. , & Sun, W. (2021). Small-sized ship detection nearshore based on lightweight active learning model with a small number of labeled data for sar imagery. Remote Sensing, 13(17), 3400.
As, I. , Pal, S. , & Basu, P. (2018). Artificial intelligence in architecture: generating conceptual design via deep learning. International Journal of Architectural Computing, 16(4), 306-327.
Aljohani, T. M. , Ebrahim, A. , & Mohammed, O. (2021). Real-time metadata-driven routing optimization for electric vehicle energy consumption minimization using deep reinforcement learning and markov chain model. Electric Power Systems Research(6), 192.
Asfahan, H. M. , Sajjad, U. , Sultan, M. , Hussain, I. , & Khan, M. U. (2021). Artificial intelligence for the prediction of the thermal performance of evaporative cooling systems. Energies, 14(13), 3946.
Ropiak, K. , & Artiemjew, P. (2020). On a hybridization of deep learning and rough set based granular computing. Algorithms, 13(3), 63.
Dong, X. , Dong, C. , Chen, Z. , Cheng, Y. , & Chen, B. (2020). Botdetector: an extreme learning machine-based internet of things botnet detection model. Transactions on Emerging Telecommunications Technologies (4), e3999.
Xia, J., Deng, D. , & Fan, D. (2020). A note on implementation methodologies of deep learning-based signal detection for conventional mimo transmitters. IEEE Transactions on Broadcasting, (99), 1-2.


