Collaborative Intelligence: Accelerating Deep Neural Network Inference via Device-Edge Synergy

تاريخ النشر

2020-09-07

دولة النشر

مصر

عدد الصفحات

التخصصات الرئيسية

تكنولوجيا المعلومات وعلم الحاسوب

الملخص EN

With the development of mobile edge computing (MEC), more and more intelligent services and applications based on deep neural networks are deployed on mobile devices to meet the diverse and personalized needs of users.

Unfortunately, deploying and inferencing deep learning models on resource-constrained devices are challenging.

The traditional cloud-based method usually runs the deep learning model on the cloud server.

Since a large amount of input data needs to be transmitted to the server through WAN, it will cause a large service latency.

This is unacceptable for most current latency-sensitive and computation-intensive applications.

In this paper, we propose Cogent, an execution framework that accelerates deep neural network inference through device-edge synergy.

In the Cogent framework, it is divided into two operation stages, including the automatic pruning and partition stage and the containerized deployment stage.

Cogent uses reinforcement learning (RL) to automatically predict pruning and partition strategies based on feedback from the hardware configuration and system conditions so that the pruned and partitioned model can better adapt to the system environment and user hardware configuration.

Then through containerized deployment to the device and the edge server to accelerate model inference, experiments show that the learning-based hardware-aware automatic pruning and partition scheme can significantly reduce the service latency, and it accelerates the overall model inference process while maintaining accuracy.

Using this method can accelerate up to 8.89× without loss of accuracy of more than 7%.

نمط استشهاد جمعية علماء النفس الأمريكية (APA)

Shan, Nanliang& Ye, Zecong& Cui, Xiaolong. 2020. Collaborative Intelligence: Accelerating Deep Neural Network Inference via Device-Edge Synergy. Security and Communication Networks،Vol. 2020, no. 2020, pp.1-10.
https://search.emarefa.net/detail/BIM-1208634

نمط استشهاد الجمعية الأمريكية للغات الحديثة (MLA)

Shan, Nanliang…[et al.]. Collaborative Intelligence: Accelerating Deep Neural Network Inference via Device-Edge Synergy. Security and Communication Networks No. 2020 (2020), pp.1-10.
https://search.emarefa.net/detail/BIM-1208634

نمط استشهاد الجمعية الطبية الأمريكية (AMA)

Shan, Nanliang& Ye, Zecong& Cui, Xiaolong. Collaborative Intelligence: Accelerating Deep Neural Network Inference via Device-Edge Synergy. Security and Communication Networks. 2020. Vol. 2020, no. 2020, pp.1-10.
https://search.emarefa.net/detail/BIM-1208634

نوع البيانات

مقالات

لغة النص

الإنجليزية

الملاحظات

Includes bibliographical references

رقم السجل

BIM-1208634

حفظتم الحفظ طباعة

قاعدة معامل التأثير والاستشهادات المرجعية العربي "ارسيف Arcif"

أضخم قاعدة بيانات عربية للاستشهادات المرجعية للمجلات العلمية المحكمة الصادرة في العالم العربي

مرصد "معرفة"
لقياس الإنتاج العلمي العربي

تقوم هذه الخدمة بالتحقق من التشابه أو الانتحال في الأبحاث والمقالات العلمية والأطروحات الجامعية والكتب والأبحاث باللغة العربية، وتحديد درجة التشابه أو أصالة الأعمال البحثية وحماية ملكيتها الفكرية. تعرف اكثر