AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (1.5 MB)
Collect
Submit Manuscript AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline

Automatic Traffic State Recognition from Videos Based on Autoencoder and k-Means Clustering

Yuanyuan Zhang1( )Yuting Wang1Bo Peng1,2Ju Tang1Jiming Xie1
College of Traffic and Transportation, Chongqing Jiaotong University, Chongqing 400074, China
Chongqing Key Lab of Traffic System & Safety in Mountain Cities, Chongqing 400074, China
Show Author Information

Abstract

An automatic video traffic state recognition method based on an autoencoder and k-means clustering is proposed to timely and effectively recognize road traffic state. First, candidate autoencoders are established through a reasonable optimization of the structural parameters on the input data dimensions, number of hidden layers, and dimensions of dimension-reduced data through cross-examination. Then, three image data sets are formed with 1500-4500 sample images. On this basis, the candidate autoencoders are trained and tested; thus, the best autoencoder AE* is proposed according to precision, recall, and F1-value. Lastly, four traffic state recognition models are constructed by combining AE* with k-means clustering, Support Vector Machine (SVM), Linear Classifier, and DNN Linear Classifier (Deep Neural Network with Linear Classifier), which are named as AE*-kmeans, AE*-SVM, AE*-Linear, and AE*-DNN_Linear, respectively. The models are trained and tested on the basis of the three image data sets. Results show that the four models' average precision in terms of precision and recall is 91.9%-92.7%, and their average recall is 91.6%-92.6%, while AE*-kmeans performs best or second to best in terms of precision and recall. With regard to the comprehensive evaluation index F1-value, AE*-kmeans achieves 92.4%, a little lower than 92.7% of AE*-SVM, and better than AE*-DNN_Linear (92.1%) and AE*-Linear (91.8%). Given that k-means is an unsupervised clustering method, compared with AE*-SVM, AE*-Linear, and AE*-DNN_Linear, AE*-kmeans can reduce the workload, such as manual data calibration and supervised training and cut down calculation cost. Meanwhile, AE*-kmeans also obtains a good traffic state recognition result. Therefore, this mechanism has high practical significance for an accurate real-time extraction of a video traffic status.

References

【1】
【1】
 
 
Journal of Highway and Transportation Research and Development (English Edition)
Pages 81-88

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Zhang Y, Wang Y, Peng B, et al. Automatic Traffic State Recognition from Videos Based on Autoencoder and k-Means Clustering. Journal of Highway and Transportation Research and Development (English Edition), 2020, 14(4): 81-88. https://doi.org/10.1061/JHTRCQ.0000757

4

Views

0

Downloads

0

Crossref

Received: 24 December 2019
Published: 01 December 2020
© The Editorial Office of Journal of Highway and Transportation Research and Development (English Edition)