| Sign up

PDF (4.1 MB)

Cite

Collect

Submit Manuscript

Open Access

Deep reinforcement learning based worker selection for distributed machine learning enhanced edge intelligence in internet of vehicles

Junyu Dong, Wenjun Wu^{^*}(), Yang Gao, Xiaoxi Wang, Pengbo Si

Faculty of Information Technology, Beijing University of Technology, Beijing 100022, China

Show Author Information

Abstract

Nowadays, Edge Information System (EIS) has received a lot of attentions. In EIS, Distributed Machine Learning (DML), which requires fewer computing resources, can implement many artificial intelligent applications efficiently. However, due to the dynamical network topology and the fluctuating transmission quality at the edge, work node selection affects the performance of DML a lot. In this paper, we focus on the Internet of Vehicles (IoV), one of the typical scenarios of EIS, and consider the DML-based High Definition (HD) mapping and intelligent driving decision model as the example. The worker selection problem is modeled as a Markov Decision Process (MDP), maximizing the DML model aggregate performance related to the timeliness of the local model, the transmission quality of model parameters uploading, and the effective sensing area of the worker. A Deep Reinforcement Learning (DRL) based solution is proposed, called the Worker Selection based on Policy Gradient (PG-WS) algorithm. The policy mapping from the system state to the worker selection action is represented by a deep neural network. The episodic simulations are built and the REINFORCE algorithm with baseline is used to train the policy network. Results show that the proposed PG-WS algorithm outperforms other comparation methods.

Keywords

edge information system internet of vehicles distributed machine learning deep reinforcement learning worker selection

References

[1]

Zhang

and K. B.

Letaief

, Mobile edge intelligence and computing for the internet of vehicles, Proc. IEEE, vol. 108, no. 2, pp. 246-261, 2020.

Crossref Google Scholar

[2]

W. C.

, H. B.

Zhou

, N.

Cheng

, F.

, W. S.

Shi

, J. Y.

Chen

, and X. M.

Shen

, Internet of vehicles in big data era, IEEE/CAA J. Autom. Sin., vol. 5, no. 1, pp. 19-35, 2018.

Crossref Google Scholar

[3]

TomTom HD map for autonomous driving extends to Japan, https://corporate.tomtom.com/news-releases/news-release-details/tomtom-hd-map-autonomous-driving-extends-japan?releaseid=1045730, 2017.

[4]

HERE introduces HD live map to show the path to highly automated driving, https://360.here.com/2016/01/05/here-introduces-hd-live-map-to-show-the-path-to-highly-automated-driving/, 2016.

[5]

P. F.

Alcantarilla

, S.

Stent

, G.

Ros

, R.

Arroyo

, and R.

Gherardi

, Street-view change detection with deconvolutional networks, Auto. Robots, vol. 42, no. 7, pp. 1301-1322, 2018.

Crossref Google Scholar

[6]

McMahan

, E.

Moore

, D.

Ramage

, S.

Hampson

, and B. A.

Arcas

, Communication-efficient learning of deep networks from decentralized data, arXiv preprint arXiv: 1602.05629, 2017.

[7]

J. M.

Chen

, X. H.

Pan

, R.

Monga

, S.

Bengio

, and R.

Jozefowicz

, Revisiting distributed synchronous SGD, arXiv preprint arXiv: 1604.00981, 2016.

[8]

S. Q.

Wang

, T.

Tuor

, T.

Salonidis

, K. K.

Leung

, C.

Makaya

, T.

, and K.

Chan

, When edge meets learning: Adaptive control for resource-constrained distributed machine learning, presented at IEEE INFOCOM 2018-IEEE Conf. Computer Communications, Honolulu, HI, USA, 2018, pp. 63-71.

Crossref

[9]

Zhang

, F. R.

, J.

Liu

, T.

Huang

, and Y. J.

Liu

, Deep reinforcement learning (DRL)-based Device-to-Device (D2D) caching with blockchain and mobile edge computing, IEEE Trans. Wireless Comm., vol. 19, no. 10, pp. 6469-6485, 2020.

Crossref Google Scholar

[10]

Gao

, W. J.

, H. X.

Nan

, Y.

Sun

, and P. B.

, Deep reinforcement learning based task scheduling in mobile Blockchain for IoT applications, presented at ICC 2020-2020 IEEE Int. Conf. Communications (ICC), Dublin, Ireland, 2020, pp. 1-7.

Crossref

[11]

, F. R.

, P. B.

, W. J.

, and Y. H.

Zhang

, Resource optimization for delay-tolerant data in blockchain-enabled iot with edge computing: A deep reinforcement learning approach, IEEE Int. Things J., vol. 7, no. 10, pp. 9399-9412, 2020.

Crossref Google Scholar

[12]

Mozaffari

, W.

Saad

, M.

Bennis

, and M.

Debbah

, Mobile unmanned aerial vehicles (UAVs) for energy-efficient internet of things communications, IEEE Trans. Wirel. Comm., vol. 16, no. 11, pp. 7574-7589, 2017.

Crossref Google Scholar

[13]

Liu

, S. W.

Liu

, and K.

Zheng

, A reinforcement learning-based resource allocation scheme for cloud robotics, IEEE Access, vol. 6, pp. 17 215-17 222, 2018.

Crossref Google Scholar

[14]

R. S.

Sutton

and A. G.

Barto

, Reinforcement Learning: An Introduction. Cambridge, MA, USA: MIT Press, 1998.

Crossref

[15]

Enhancement of 3GPP Support for V2X Scenarios, 3GPP TS 22.186, 2019.

Intelligent and Converged Networks

Volume 1 Issue 3,
December 2020

Pages 234-242

DOI: 10.23919/ICN.2020.0015

Cite this article:

Dong J, Wu W, Gao Y, et al. Deep reinforcement learning based worker selection for distributed machine learning enhanced edge intelligence in internet of vehicles. Intelligent and Converged Networks, 2020, 1(3): 234-242. https://doi.org/10.23919/ICN.2020.0015