{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,31]],"date-time":"2026-07-31T21:29:39Z","timestamp":1785533379090,"version":"3.56.0"},"reference-count":30,"publisher":"MDPI AG","issue":"12","license":[{"start":{"date-parts":[[2018,12,4]],"date-time":"2018-12-04T00:00:00Z","timestamp":1543881600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"Chongqing Research Program of Basic Science and Frontier Technology","award":["cstc2017jcyjB0305"],"award-info":[{"award-number":["cstc2017jcyjB0305"]}]},{"name":"National Key R&amp;D Program of China","award":["2017YFB0802400"],"award-info":[{"award-number":["2017YFB0802400"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Sensors"],"abstract":"<jats:p>Vehicle detection is one of the important applications of object detection in intelligent transportation systems. It aims to extract specific vehicle-type information from pictures or videos containing vehicles. To solve the problems of existing vehicle detection, such as the lack of vehicle-type recognition, low detection accuracy, and slow speed, a new vehicle detection model YOLOv2_Vehicle based on YOLOv2 is proposed in this paper. The k-means++ clustering algorithm was used to cluster the vehicle bounding boxes on the training dataset, and six anchor boxes with different sizes were selected. Considering that the different scales of the vehicles may influence the vehicle detection model, normalization was applied to improve the loss calculation method for length and width of bounding boxes. To improve the feature extraction ability of the network, the multi-layer feature fusion strategy was adopted, and the repeated convolution layers in high layers were removed. The experimental results on the Beijing Institute of Technology (BIT)-Vehicle validation dataset demonstrated that the mean Average Precision (mAP) could reach 94.78%. The proposed model also showed excellent generalization ability on the CompCars test dataset, where the \u201cvehicle face\u201d is quite different from the training dataset. With the comparison experiments, it was proven that the proposed method is effective for vehicle detection. In addition, with network visualization, the proposed model showed excellent feature extraction ability.<\/jats:p>","DOI":"10.3390\/s18124272","type":"journal-article","created":{"date-parts":[[2018,12,4]],"date-time":"2018-12-04T11:56:18Z","timestamp":1543924578000},"page":"4272","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":191,"title":["An Improved YOLOv2 for Vehicle Detection"],"prefix":"10.3390","volume":"18","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-8703-7310","authenticated-orcid":false,"given":"Jun","family":"Sang","sequence":"first","affiliation":[{"name":"Key Laboratory of Dependable Service Computing in Cyber Physical Society of Ministry of Education, Chongqing University, Chongqing 40004, China"},{"name":"School of Big Data &amp; Software Engineering, Chongqing University, Chongqing 401331, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-7916-4259","authenticated-orcid":false,"given":"Zhongyuan","family":"Wu","sequence":"additional","affiliation":[{"name":"Key Laboratory of Dependable Service Computing in Cyber Physical Society of Ministry of Education, Chongqing University, Chongqing 40004, China"},{"name":"School of Big Data &amp; Software Engineering, Chongqing University, Chongqing 401331, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Pei","family":"Guo","sequence":"additional","affiliation":[{"name":"Key Laboratory of Dependable Service Computing in Cyber Physical Society of Ministry of Education, Chongqing University, Chongqing 40004, China"},{"name":"School of Big Data &amp; Software Engineering, Chongqing University, Chongqing 401331, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Haibo","family":"Hu","sequence":"additional","affiliation":[{"name":"Key Laboratory of Dependable Service Computing in Cyber Physical Society of Ministry of Education, Chongqing University, Chongqing 40004, China"},{"name":"School of Big Data &amp; Software Engineering, Chongqing University, Chongqing 401331, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Hong","family":"Xiang","sequence":"additional","affiliation":[{"name":"Key Laboratory of Dependable Service Computing in Cyber Physical Society of Ministry of Education, Chongqing University, Chongqing 40004, China"},{"name":"School of Big Data &amp; Software Engineering, Chongqing University, Chongqing 401331, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Qian","family":"Zhang","sequence":"additional","affiliation":[{"name":"Key Laboratory of Dependable Service Computing in Cyber Physical Society of Ministry of Education, Chongqing University, Chongqing 40004, China"},{"name":"School of Big Data &amp; Software Engineering, Chongqing University, Chongqing 401331, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Bin","family":"Cai","sequence":"additional","affiliation":[{"name":"Key Laboratory of Dependable Service Computing in Cyber Physical Society of Ministry of Education, Chongqing University, Chongqing 40004, China"},{"name":"School of Big Data &amp; Software Engineering, Chongqing University, Chongqing 401331, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"1968","published-online":{"date-parts":[[2018,12,4]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","unstructured":"Cao, X., Wu, C., Yan, P., and Li, X. (2011, January 11\u201314). Linear SVM classification using boosting HOG features for vehicle detection in low-altitude airborne videos. Proceedings of the 2011 IEEE International Conference Image Processing (ICIP), Brussels, Belgium.","DOI":"10.1109\/ICIP.2011.6116132"},{"key":"ref_2","doi-asserted-by":"crossref","unstructured":"Guo, E., Bai, L., Zhang, Y., and Han, J. (2017, January 13\u201315). Vehicle Detection Based on Superpixel and Improved HOG in Aerial Images. Proceedings of the International Conference on Image and Graphics, Shanghai, China.","DOI":"10.1007\/978-3-319-71607-7_32"},{"key":"ref_3","doi-asserted-by":"crossref","unstructured":"Laopracha, N., and Sunat, K. (2017, January 21\u201323). Comparative Study of Computational Time that HOG-Based Features Used for Vehicle Detection. Proceedings of the International Conference on Computing and Information Technology, Helsinki, Finland.","DOI":"10.1007\/978-3-319-60663-7_26"},{"key":"ref_4","unstructured":"Pan, C., Sun, M., and Yan, Z. (2016, January 13\u201315). The Study on Vehicle Detection Based on DPM in Traffic Scenes. Proceedings of the International Conference on Frontier Computing, Tokyo, Japan."},{"key":"ref_5","doi-asserted-by":"crossref","first-page":"436","DOI":"10.1038\/nature14539","article-title":"Deep learning","volume":"521","author":"LeCun","year":"2015","journal-title":"Nature"},{"key":"ref_6","doi-asserted-by":"crossref","unstructured":"He, K., Zhang, X., Ren, S., and Sun, J. (2016, January 27\u201330). Deep residual learning for image recognition. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Las Vegas, NV, USA.","DOI":"10.1109\/CVPR.2016.90"},{"key":"ref_7","doi-asserted-by":"crossref","unstructured":"Huang, G., Liu, Z., Van Der Maaten, L., and Weinberger, K.Q. (2017, January 22\u201325). Densely Connected Convolutional Networks. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.243"},{"key":"ref_8","doi-asserted-by":"crossref","unstructured":"Pyo, J., Bang, J., and Jeong, Y. (2016, January 23\u201326). Front collision warning based on vehicle detection using CNN. Proceedings of the International SoC Design Conference (ISOCC), Jeju, Korea.","DOI":"10.1109\/ISOCC.2016.7799842"},{"key":"ref_9","doi-asserted-by":"crossref","first-page":"5817","DOI":"10.1007\/s11042-015-2520-x","article-title":"Vehicle detection and recognition for intelligent traffic surveillance system","volume":"76","author":"Tang","year":"2017","journal-title":"Multimed. Tools Appl."},{"key":"ref_10","doi-asserted-by":"crossref","unstructured":"Gao, Y., Guo, S., Huang, K., Chen, J., Gong, Q., Zou, Y., Bai, T., and Overett, G. (2017, January 11\u201314). Scale optimization for full-image-CNN vehicle detection. Proceedings of the IEEE Intelligent Vehicles Symposium (IV), Los Angeles, CA, USA.","DOI":"10.1109\/IVS.2017.7995812"},{"key":"ref_11","doi-asserted-by":"crossref","unstructured":"Huttunen, H., Yancheshmeh, F.S., and Chen, K. (2016, January 19\u201322). Car type recognition with deep neural networks. Proceedings of the IEEE Intelligent Vehicles Symposium (IV), Gothenburg, Sweden.","DOI":"10.1109\/IVS.2016.7535529"},{"key":"ref_12","doi-asserted-by":"crossref","unstructured":"Dong, Z., Pei, M., and He, Y. (2014, January 24\u201328). Vehicle type classification using unsupervised convolution neural network. Proceedings of the 2014 IEEE International Conference on Pattern Recognition, Stockholm, Sweden.","DOI":"10.1109\/ICPR.2014.39"},{"key":"ref_13","doi-asserted-by":"crossref","unstructured":"Girshick, R., Donahue, J., Darrell, T., and Malik, J. (2014, January 24\u201327). Rich feature hierarchies for accurate object detection and semantic segmentation. Proceedings of the 2014 IEEE Conference on Computer Vision and Pattern Recognition, Columbus, OH, USA.","DOI":"10.1109\/CVPR.2014.81"},{"key":"ref_14","doi-asserted-by":"crossref","unstructured":"He, K., Zhang, X., Ren, S., and Sun, J. (2014, January 6\u201312). Spatial pyramid pooling in deep convolutional networks for visual recognition. Proceedings of the 2014 IEEE International Conference of European Conference on Computer Vision, Zurich, Switzerland.","DOI":"10.1007\/978-3-319-10578-9_23"},{"key":"ref_15","doi-asserted-by":"crossref","unstructured":"Girshick, R. (2015, January 7\u201313). Fast R-CNN. Proceedings of the 2015 IEEE International Conference on Computer Vision (ICCV), Santiago, Chile.","DOI":"10.1109\/ICCV.2015.169"},{"key":"ref_16","doi-asserted-by":"crossref","unstructured":"Ren, S., He, K., Girshick, R., and Sun, J. (arXiv, 2016). Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks, arXiv.","DOI":"10.1109\/TPAMI.2016.2577031"},{"key":"ref_17","unstructured":"Dai, J., Li, Y., He, K., and Sun, J. (2016, January 5\u20138). R-FCN: Object detection via region-based fully convolutional networks. Proceedings of the 2016 IEEE International Conference of Advances in neural information processing systems, Barcelona, Spain."},{"key":"ref_18","doi-asserted-by":"crossref","unstructured":"Konoplich, G.V., Putin, E.O., and Filchenkov, A.A. (2016, January 25\u201327). Application of deep learning to the problem of vehicle detection in UAV images. Proceedings of the 2016 XIX IEEE International Conference on Soft Computing and Measurements (SCM), St. Petersburg, Russia.","DOI":"10.1109\/SCM.2016.7519666"},{"key":"ref_19","doi-asserted-by":"crossref","unstructured":"Cai, Z., Fan, Q., Feris, R.S., and Vasconcelos, N. (2016, January 8\u201316). A unified multi-scale deep convolutional neural network for fast object detection. Proceedings of the 2016 European Conference on Computer Vision, Amsterdam, The Netherlands.","DOI":"10.1007\/978-3-319-46493-0_22"},{"key":"ref_20","doi-asserted-by":"crossref","unstructured":"Azam, S., Rafique, A., and Jeon, M. (2016, January 27\u201329). Vehicle pose detection using region based convolutional neural network. Proceedings of the International Conference on Control, Automation and Information Sciences (ICCAIS), Ansan, Korea.","DOI":"10.1109\/ICCAIS.2016.7822459"},{"key":"ref_21","doi-asserted-by":"crossref","unstructured":"Tang, T., Zhou, S., Deng, Z., Zou, H., and Lei, L. (2017). Vehicle detection in aerial images based on region convolutional neural networks and hard negative example mining. Sensors, 17.","DOI":"10.3390\/s17020336"},{"key":"ref_22","first-page":"32","article-title":"Vehicle detection based on faster-RCNN","volume":"40","author":"Sang","year":"2017","journal-title":"J. Chongqing Univ. (Nat. Sci. Ed.)"},{"key":"ref_23","doi-asserted-by":"crossref","unstructured":"Redmon, J., Divvala, S., Girshick, R., and Farhadi, A. (2016, January 27\u201330). You only look once: Unified, real-time object detection. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Las Vegas, NV, USA.","DOI":"10.1109\/CVPR.2016.91"},{"key":"ref_24","doi-asserted-by":"crossref","unstructured":"Redmon, J., and Farhadi, A. (2017, January 21\u201326). YOLO9000: Better, Faster, Stronger. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.690"},{"key":"ref_25","unstructured":"Arthur, D., and Vassilvitskii, S. (2007, January 7\u20139). k-means++: The advantages of careful seeding. Proceedings of the eighteenth annual ACM-SIAM symposium on Discrete algorithms, New Orleans, LA, USA."},{"key":"ref_26","doi-asserted-by":"crossref","unstructured":"Neubeck, A., and Van Gool, L. (2006, January 20\u201324). Efficient non-maximum suppression. Proceedings of the International Conference on Pattern Recognition (ICPR), Hong Kong, China.","DOI":"10.1109\/ICPR.2006.479"},{"key":"ref_27","unstructured":"Ioffe, S., and Szegedy, C. (2005, January 6\u201311). Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Shift. Proceedings of the International Conference on Machine Learning, Lille, France."},{"key":"ref_28","doi-asserted-by":"crossref","first-page":"2247","DOI":"10.1109\/TITS.2015.2402438","article-title":"Vehicle type classification using a semisupervised convolutional neural network","volume":"16","author":"Dong","year":"2015","journal-title":"IEEE Trans. Intel. Transp. Syst."},{"key":"ref_29","doi-asserted-by":"crossref","unstructured":"Yang, L., Luo, P., Change Loy, C., and Tang, X. (2015, January 7\u201312). A large-scale car dataset for fine-grained categorization and verification. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Boston, MA, USA.","DOI":"10.1109\/CVPR.2015.7299023"},{"key":"ref_30","doi-asserted-by":"crossref","unstructured":"Zeiler, M.D., and Fergus, R. (2014, January 6\u201312). Visualizing and understanding convolutional networks. Proceedings of the European Conference on Computer Vision (ECCV), Zurich, Switzerland.","DOI":"10.1007\/978-3-319-10590-1_53"}],"container-title":["Sensors"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/1424-8220\/18\/12\/4272\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,11]],"date-time":"2025-10-11T15:31:10Z","timestamp":1760196670000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/1424-8220\/18\/12\/4272"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2018,12,4]]},"references-count":30,"journal-issue":{"issue":"12","published-online":{"date-parts":[[2018,12]]}},"alternative-id":["s18124272"],"URL":"https:\/\/doi.org\/10.3390\/s18124272","relation":{},"ISSN":["1424-8220"],"issn-type":[{"value":"1424-8220","type":"electronic"}],"subject":[],"published":{"date-parts":[[2018,12,4]]}}}