Inventi Impact: Soft Computing

Articles

Inventi:esc/33414/20

Collaborative Intelligence: Accelerating Deep Neural Network Inference via Device-Edge Synergy

Research 2021 : January - March

Nanliang Shan, Zecong Ye, Xiaolong Cui

With the development of mobile edge computing (MEC), more and more intelligent services and applications based on deep\nneural networks are deployed on mobile devices to meet the diverse and personalized needs of users. Unfortunately, deploying and\ninferencing deep learning models on resource-constrained devices are challenging. The traditional cloud-based method usually\nruns the deep learning model on the cloud server. Since a large amount of input data needs to be transmitted to the server through\nWAN, it will cause a large service latency. This is unacceptable for most current latency-sensitive and computation-intensive\napplications. In this paper, we propose Cogent, an execution framework that accelerates deep neural network inference through\ndevice-edge synergy. In the Cogent framework, it is divided into two operation stages, including the automatic pruning and\npartition stage and the containerized deployment stage. Cogent uses reinforcement learning (RL) to automatically predict pruning\nand partition strategies based on feedback from the hardware configuration and system conditions so that the pruned and\npartitioned model can better adapt to the system environment and user hardware configuration. Then through containerized\ndeployment to the device and the edge server to accelerate model inference, experiments show that the learning-based hardwareaware\nautomatic pruning and partition scheme can significantly reduce the service latency, and it accelerates the overall model\ninference process while maintaining accuracy.

How to Cite this Article
Attribution/ CC Compliant Citation: Shan, Nanliang, Zecong Ye,\nand Xiaolong Cui. \"Collaborative Intelligence: Accelerating Deep\nNeural Network Inference via Device-Edge Synergy.\" Security and\nCommunication Networks 2020 (2020).\nhttps://doi.org/10.1155/2020/8831341\nhttps://creativecommons.org/licenses/by/4.0/\nSome formatting elements, header, footer, logos, dates and\npagination were modified while adapting this article.
Download Full Text

Call Us: +4 (800) 888-0008

Inventi Impact: Soft Computing

Articles

Inventi:esc/33414/20

Collaborative Intelligence: Accelerating Deep Neural Network Inference via Device-Edge Synergy

How to Cite this Article

Links

Contact Us