{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,8,21]],"date-time":"2026-08-21T13:20:36Z","timestamp":1787318436814,"version":"build-2736575974"},"reference-count":171,"publisher":"Association for Computing Machinery (ACM)","issue":"2","license":[{"start":{"date-parts":[[2021,3,5]],"date-time":"2021-03-05T00:00:00Z","timestamp":1614902400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"NSERC-SPG, NSERC-DISCOVERY, Canada Research Chairs Program, and NSERCCREATE TRANSIT Funds"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Comput. Surv."],"published-print":{"date-parts":[[2022,3,31]]},"abstract":"<jats:p>The recent boom of autonomous driving nowadays has made object detection in traffic scenes a hot topic of research. Designed to classify and locate instances in the image, this is a basic but challenging task in the computer vision field. With its powerful feature extraction abilities, which are vital for object detection, deep learning has expanded its application areas to this field during the past several years and thus achieved breakthroughs. However, even with such powerful approaches, traffic scenarios have their own specific challenges, such as real-time detection, changeable weather, and complex lighting conditions. This survey is dedicated to summarizing research and papers on applying deep learning to the transportation environment in recent years. More than 100 research papers are covered, and different aspects such as key generic object detection frameworks, categorized object detection applications in traffic scenario, evaluation metrics, and classified datasets are included. Some open research fields are also provided. We believe that it is the first survey focusing on deep learning-based object detection in traffic scenario.<\/jats:p>","DOI":"10.1145\/3434398","type":"journal-article","created":{"date-parts":[[2021,3,5]],"date-time":"2021-03-05T23:09:57Z","timestamp":1614985797000},"page":"1-35","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":106,"title":["Object Detection Using Deep Learning Methods in Traffic Scenarios"],"prefix":"10.1145","volume":"54","author":[{"given":"Azzedine","family":"Boukerche","sequence":"first","affiliation":[{"name":"University of Ottawa, Ottawa, ON, Canada"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Zhijun","family":"Hou","sequence":"additional","affiliation":[{"name":"University of Ottawa, Ottawa, ON, Canada"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2021,3,5]]},"reference":[{"key":"e_1_2_1_1_1","unstructured":"La Route Automatis\u00e9e. 2019. Traffic Lights Recognition (TLR) public benchmarks. Retrieved from http:\/\/www.lara.prd.fr\/benchmarks\/trafficlightsrecognition.  La Route Automatis\u00e9e. 2019. Traffic Lights Recognition (TLR) public benchmarks. Retrieved from http:\/\/www.lara.prd.fr\/benchmarks\/trafficlightsrecognition."},{"key":"e_1_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1109\/ITSC.2018.8569522"},{"key":"e_1_2_1_3_1","volume-title":"Proceedings of the IEEE International Conference on Robotics and Automation (ICRA\u201917)","author":"Behrendt Karsten","unstructured":"Karsten Behrendt and Libor Novak . [n.d.]. A deep learning approach to traffic lights: Detection, tracking, and classification . In Proceedings of the IEEE International Conference on Robotics and Automation (ICRA\u201917) . IEEE. Karsten Behrendt and Libor Novak. [n.d.]. A deep learning approach to traffic lights: Detection, tracking, and classification. In Proceedings of the IEEE International Conference on Robotics and Automation (ICRA\u201917). IEEE."},{"key":"e_1_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2017.2699184"},{"key":"e_1_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.236"},{"key":"e_1_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2017.2706685"},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.691"},{"key":"e_1_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-73603-7_27"},{"key":"e_1_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2017.2762591"},{"key":"e_1_2_1_10_1","unstructured":"Embedded Computing Lab. 2019. WPI traffic light dataset. Retrieved from http:\/\/computing.wpi.edu\/dataset.html.  Embedded Computing Lab. 2019. WPI traffic light dataset. Retrieved from http:\/\/computing.wpi.edu\/dataset.html."},{"key":"e_1_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.350"},{"key":"e_1_2_1_12_1","unstructured":"Marius Cordts Mohamed Omran Sebastian Ramos Timo Rehfeld Markus Enzweiler Rodrigo Benenson Uwe Franke Stefan Roth and Bernt Schiele. 2019. Cityscapes dataset. Retrieved from https:\/\/www.cityscapes-dataset.com\/.  Marius Cordts Mohamed Omran Sebastian Ramos Timo Rehfeld Markus Enzweiler Rodrigo Benenson Uwe Franke Stefan Roth and Bernt Schiele. 2019. Cityscapes dataset. Retrieved from https:\/\/www.cityscapes-dataset.com\/."},{"key":"e_1_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIT.1967.1053964"},{"key":"e_1_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1111\/j.2517-6161.1958.tb00292.x"},{"key":"e_1_2_1_15_1","unstructured":"Autti CrowdAI. 2017. Udacity labeled dataset. Retrieved from https:\/\/github.com\/udacity\/self-driving-car\/tree\/master\/annotations.  Autti CrowdAI. 2017. Udacity labeled dataset. Retrieved from https:\/\/github.com\/udacity\/self-driving-car\/tree\/master\/annotations."},{"key":"e_1_2_1_16_1","unstructured":"Yaodong Cui Ren Chen Wenbo Chu Long Chen Daxin Tian and Dongpu Cao. 2020. Deep learning for image and point cloud fusion in autonomous driving: A review. Retrieved from https:\/\/Arxiv:2004.05224.  Yaodong Cui Ren Chen Wenbo Chu Long Chen Daxin Tian and Dongpu Cao. 2020. Deep learning for image and point cloud fusion in autonomous driving: A review. Retrieved from https:\/\/Arxiv:2004.05224."},{"key":"e_1_2_1_17_1","volume-title":"R-fcn: Object detection via region-based fully convolutional networks. In Advances in Neural Information Processing Systems","author":"Dai Jifeng","year":"2016","unstructured":"Jifeng Dai , Yi Li , Kaiming He , and Jian Sun . 2016 . R-fcn: Object detection via region-based fully convolutional networks. In Advances in Neural Information Processing Systems . MIT Press , 379--387. Jifeng Dai, Yi Li, Kaiming He, and Jian Sun. 2016. R-fcn: Object detection via region-based fully convolutional networks. In Advances in Neural Information Processing Systems. MIT Press, 379--387."},{"key":"e_1_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2005.177"},{"key":"e_1_2_1_19_1","unstructured":"Navneet Dalal and Bill Triggs. 2019. INRIA Person Dataset. Retrieved from http:\/\/pascal.inrialpes.fr\/data\/human\/.  Navneet Dalal and Bill Triggs. 2019. INRIA Person Dataset. Retrieved from http:\/\/pascal.inrialpes.fr\/data\/human\/."},{"key":"e_1_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.1145\/1143844.1143874"},{"key":"e_1_2_1_21_1","unstructured":"Grupo de Tratamiento de Imagenes (GTI). 2012. GTI vehicle image database. Retrieved from http:\/\/www.gti.ssr.upm.es\/data\/Vehicle_database.html.  Grupo de Tratamiento de Imagenes (GTI). 2012. GTI vehicle image database. Retrieved from http:\/\/www.gti.ssr.upm.es\/data\/Vehicle_database.html."},{"key":"e_1_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-23222-5_25"},{"key":"e_1_2_1_23_1","volume-title":"Proceedings of the International Conference on Computer Vision & Pattern Recognition (CVPR\u201909)","author":"Doll\u00e1r P.","unstructured":"P. Doll\u00e1r , C. Wojek , B. Schiele , and P. Perona . 2009. Pedestrian detection: A benchmark . In Proceedings of the International Conference on Computer Vision & Pattern Recognition (CVPR\u201909) . P. Doll\u00e1r, C. Wojek, B. Schiele, and P. Perona. 2009. Pedestrian detection: A benchmark. In Proceedings of the International Conference on Computer Vision & Pattern Recognition (CVPR\u201909)."},{"key":"e_1_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2011.155"},{"key":"e_1_2_1_25_1","volume-title":"Pedestrian detection: An evaluation of the state of the art. PAMI 34","author":"Doll\u00e1r Piotr","year":"2012","unstructured":"Piotr Doll\u00e1r , Christian Wojek , Bernt Schiele , and Pietro Perona . 2012. Pedestrian detection: An evaluation of the state of the art. PAMI 34 ( 2012 ). Piotr Doll\u00e1r, Christian Wojek, Bernt Schiele, and Pietro Perona. 2012. Pedestrian detection: An evaluation of the state of the art. PAMI 34 (2012)."},{"key":"e_1_2_1_26_1","unstructured":"Piotr Doll\u00e1r Christian Wojek Bernt Schiele and Pietro Perona. 2019. Caltech Pedestrian Detection Benchmark. Retrieved from http:\/\/www.vision.caltech.edu\/Image_Datasets\/CaltechPedestrians\/.  Piotr Doll\u00e1r Christian Wojek Bernt Schiele and Pietro Perona. 2019. Caltech Pedestrian Detection Benchmark. Retrieved from http:\/\/www.vision.caltech.edu\/Image_Datasets\/CaltechPedestrians\/."},{"key":"e_1_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2015.316"},{"key":"e_1_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1109\/WACV.2017.111"},{"key":"e_1_2_1_29_1","doi-asserted-by":"crossref","unstructured":"Kaiwen Duan Song Bai Lingxi Xie Honggang Qi Qingming Huang and Qi Tian. 2019. CenterNet: Keypoint triplets for object detection. Retrieved from https:\/\/Arxiv:1904.08189.  Kaiwen Duan Song Bai Lingxi Xie Honggang Qi Qingming Huang and Qi Tian. 2019. CenterNet: Keypoint triplets for object detection. Retrieved from https:\/\/Arxiv:1904.08189.","DOI":"10.1109\/ICCV.2019.00667"},{"key":"e_1_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2008.260"},{"key":"e_1_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2007.4409092"},{"key":"e_1_2_1_32_1","unstructured":"Andreas Ess Bastian Leibe and Luc Van Gool. 2019. Robust Multi-Person Tracking from Mobile Platforms. Retrieved from https:\/\/data.vision.ee.ethz.ch\/cvl\/aess\/dataset\/.  Andreas Ess Bastian Leibe and Luc Van Gool. 2019. Robust Multi-Person Tracking from Mobile Platforms. Retrieved from https:\/\/data.vision.ee.ethz.ch\/cvl\/aess\/dataset\/."},{"key":"e_1_2_1_33_1","unstructured":"Mark Everingham Luc van Gool Chris Williams John Winn and Andrew Zisserman. 2005. Visual Object Classes Challenge 2012 (VOC2012). Retrieved from http:\/\/host.robots.ox.ac.uk\/pascal\/VOC\/voc2012\/.  Mark Everingham Luc van Gool Chris Williams John Winn and Andrew Zisserman. 2005. Visual Object Classes Challenge 2012 (VOC2012). Retrieved from http:\/\/host.robots.ox.ac.uk\/pascal\/VOC\/voc2012\/."},{"key":"e_1_2_1_34_1","doi-asserted-by":"publisher","DOI":"10.1109\/IVS.2016.7535375"},{"key":"e_1_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2008.4587597"},{"key":"e_1_2_1_36_1","unstructured":"Heidelberg Collaboratory for Image Processing. 2019. Bosch Small Traffic Lights Dataset. Retrieved from https:\/\/hci.iwr.uni-heidelberg.de\/node\/6132.  Heidelberg Collaboratory for Image Processing. 2019. Bosch Small Traffic Lights Dataset. Retrieved from https:\/\/hci.iwr.uni-heidelberg.de\/node\/6132."},{"key":"e_1_2_1_37_1","unstructured":"The Laboratory for Intelligent and Safe Automobiles. 2010. Vehicle Detection Dataset. Retrieved from http:\/\/cvrr.ucsd.edu\/LISA\/vehicledetection.html.  The Laboratory for Intelligent and Safe Automobiles. 2010. Vehicle Detection Dataset. Retrieved from http:\/\/cvrr.ucsd.edu\/LISA\/vehicledetection.html."},{"key":"e_1_2_1_38_1","unstructured":"Vision for Intelligent Vehicles and Applications. 2006. VIVA traffic light detection benchmark. Retrieved from http:\/\/cvrr.ucsd.edu\/vivachallenge\/index.php\/traffic-light\/traffic-light-detection\/.  Vision for Intelligent Vehicles and Applications. 2006. VIVA traffic light detection benchmark. Retrieved from http:\/\/cvrr.ucsd.edu\/vivachallenge\/index.php\/traffic-light\/traffic-light-detection\/."},{"key":"e_1_2_1_39_1","doi-asserted-by":"publisher","DOI":"10.1006\/jcss.1997.1504"},{"key":"e_1_2_1_40_1","doi-asserted-by":"publisher","DOI":"10.1109\/ITSC.2013.6728473"},{"key":"e_1_2_1_41_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICWAPR.2010.5576425"},{"key":"e_1_2_1_42_1","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 6926--6935","author":"Gao Mingfei","unstructured":"Mingfei Gao , Ruichi Yu , Ang Li , Vlad I. Morariu , and Larry S. Davis . 2018. Dynamic zoom-in network for fast object detection in large images . In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 6926--6935 . Mingfei Gao, Ruichi Yu, Ang Li, Vlad I. Morariu, and Larry S. Davis. 2018. Dynamic zoom-in network for fast object detection in large images. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 6926--6935."},{"key":"e_1_2_1_43_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2012.6248074"},{"key":"e_1_2_1_44_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2009.122"},{"key":"e_1_2_1_45_1","volume-title":"Fast R-CNN. In Proceedings of the IEEE International Conference on Computer Vision. 1440--1448","author":"Girshick Ross","year":"2015","unstructured":"Ross Girshick . 2015 . Fast R-CNN. In Proceedings of the IEEE International Conference on Computer Vision. 1440--1448 . Ross Girshick. 2015. Fast R-CNN. In Proceedings of the IEEE International Conference on Computer Vision. 1440--1448."},{"key":"e_1_2_1_46_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2014.81"},{"key":"e_1_2_1_47_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2015.2437384"},{"key":"e_1_2_1_48_1","doi-asserted-by":"publisher","DOI":"10.1109\/TITS.2012.2208909"},{"key":"e_1_2_1_49_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2015.2389824"},{"key":"e_1_2_1_50_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.90"},{"key":"e_1_2_1_51_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-46493-0_38"},{"key":"e_1_2_1_52_1","first-page":"1","article-title":"Pedestrian detection at night using deep neural networks and saliency maps","volume":"2018","author":"Heo Duyoung","year":"2018","unstructured":"Duyoung Heo , Eunju Lee , and Byoung Chul Ko . 2018 . Pedestrian detection at night using deep neural networks and saliency maps . Electron. Imag. 2018 , 17 (2018), 1 -- 9 . Duyoung Heo, Eunju Lee, and Byoung Chul Ko. 2018. Pedestrian detection at night using deep neural networks and saliency maps. Electron. Imag. 2018, 17 (2018), 1--9.","journal-title":"Electron. Imag."},{"key":"e_1_2_1_53_1","unstructured":"Congrui Hetang Hongwei Qin Shaohui Liu and Junjie Yan. 2017. Impression network for video object detection. Retrieved from https:\/\/Arxiv:1712.05896.  Congrui Hetang Hongwei Qin Shaohui Liu and Junjie Yan. 2017. Impression network for video object detection. Retrieved from https:\/\/Arxiv:1712.05896."},{"key":"e_1_2_1_54_1","doi-asserted-by":"publisher","DOI":"10.1162\/neco.1997.9.8.1735"},{"key":"e_1_2_1_55_1","doi-asserted-by":"publisher","DOI":"10.1109\/IJCNN.2013.6706807"},{"key":"e_1_2_1_56_1","unstructured":"Sebastian Houben Johannes Stallkamp Jan Salmen Marc Schlipsing and Christian Igel. 2013. German Traffic Sign Detection Benchmark. Retrieved from http:\/\/benchmark.ini.rub.de\/?section&equals;gtsdb&subsection&equals;&equals;news.  Sebastian Houben Johannes Stallkamp Jan Salmen Marc Schlipsing and Christian Igel. 2013. German Traffic Sign Detection Benchmark. Retrieved from http:\/\/benchmark.ini.rub.de\/?section&equals;gtsdb&subsection&equals;&equals;news."},{"key":"e_1_2_1_58_1","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 4700--4708","author":"Huang Gao","unstructured":"Gao Huang , Zhuang Liu , Laurens Van Der Maaten , and Kilian Q. Weinberger . 2017. Densely connected convolutional networks . In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 4700--4708 . Gao Huang, Zhuang Liu, Laurens Van Der Maaten, and Kilian Q. Weinberger. 2017. Densely connected convolutional networks. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 4700--4708."},{"key":"e_1_2_1_59_1","doi-asserted-by":"publisher","DOI":"10.1109\/CAC.2018.8623093"},{"key":"e_1_2_1_60_1","unstructured":"ILSVRC. 2019. ImageNet Large Scale Visual Recognition Challenge. Retrieved from http:\/\/www.image-net.org\/challenges\/LSVRC\/.  ILSVRC. 2019. ImageNet Large Scale Visual Recognition Challenge. Retrieved from http:\/\/www.image-net.org\/challenges\/LSVRC\/."},{"key":"e_1_2_1_61_1","unstructured":"Sergey Ioffe and Christian Szegedy. 2015. Batch normalization: Accelerating deep network training by reducing internal covariate shift. Retrieved from https:\/\/Arxiv:1502.03167.  Sergey Ioffe and Christian Szegedy. 2015. Batch normalization: Accelerating deep network training by reducing internal covariate shift. Retrieved from https:\/\/Arxiv:1502.03167."},{"key":"e_1_2_1_62_1","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops. 9--15","author":"Jensen Morten B.","unstructured":"Morten B. Jensen , Kamal Nasrollahi , and Thomas B. Moeslund . 2017. Evaluating state-of-the-art object detector on challenging traffic light data . In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops. 9--15 . Morten B. Jensen, Kamal Nasrollahi, and Thomas B. Moeslund. 2017. Evaluating state-of-the-art object detector on challenging traffic light data. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops. 9--15."},{"key":"e_1_2_1_63_1","doi-asserted-by":"publisher","DOI":"10.1109\/TITS.2015.2509509"},{"key":"e_1_2_1_64_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2019.2939201"},{"key":"e_1_2_1_65_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICSIPA.2017.8120665"},{"key":"e_1_2_1_66_1","unstructured":"Narendra Kumar Kamila. 2015. Handbook of Research on Emerging Perspectives in Intelligent Pattern Recognition Analysis and Image Processing. IGI Global.  Narendra Kumar Kamila. 2015. Handbook of Research on Emerging Perspectives in Intelligent Pattern Recognition Analysis and Image Processing. IGI Global."},{"key":"e_1_2_1_67_1","doi-asserted-by":"publisher","DOI":"10.1109\/BigData.2017.8258427"},{"key":"e_1_2_1_68_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCE-Asia.2016.7804765"},{"key":"e_1_2_1_69_1","unstructured":"Tao Kong Fuchun Sun Huaping Liu Yuning Jiang and Jianbo Shi. 2019. FoveaBox: Beyond anchor-based object detector. Retrieved from https:\/\/Arxiv:1904.03797.  Tao Kong Fuchun Sun Huaping Liu Yuning Jiang and Jianbo Shi. 2019. FoveaBox: Beyond anchor-based object detector. Retrieved from https:\/\/Arxiv:1904.03797."},{"key":"e_1_2_1_70_1","volume-title":"Advances in Neural Information Processing Systems","author":"Kr\u00e4henb\u00fchl Philipp","unstructured":"Philipp Kr\u00e4henb\u00fchl and Vladlen Koltun . 2011. Efficient inference in fully connected crfs with gaussian edge potentials . In Advances in Neural Information Processing Systems . MIT Press , 109--117. Philipp Kr\u00e4henb\u00fchl and Vladlen Koltun. 2011. Efficient inference in fully connected crfs with gaussian edge potentials. In Advances in Neural Information Processing Systems. MIT Press, 109--117."},{"key":"e_1_2_1_71_1","volume-title":"Hinton","author":"Krizhevsky Alex","year":"2012","unstructured":"Alex Krizhevsky , Ilya Sutskever , and Geoffrey E . Hinton . 2012 . Imagenet classification with deep convolutional neural networks. In Advances in Neural Information Processing Systems. MIT Press , 1097--1105. Alex Krizhevsky, Ilya Sutskever, and Geoffrey E. Hinton. 2012. Imagenet classification with deep convolutional neural networks. In Advances in Neural Information Processing Systems. MIT Press, 1097--1105."},{"key":"e_1_2_1_72_1","volume-title":"Proceedings of the IEEE\/RSJ International Conference on Intelligent Robots and Systems (IROS\u201918)","author":"Ku Jason","unstructured":"Jason Ku , Melissa Mozifian , Jungwook Lee , Ali Harakeh , and Steven L. Waslander . 2018. Joint 3D proposal generation and object detection from view aggregation . In Proceedings of the IEEE\/RSJ International Conference on Intelligent Robots and Systems (IROS\u201918) . IEEE, 1--8. Jason Ku, Melissa Mozifian, Jungwook Lee, Ali Harakeh, and Steven L. Waslander. 2018. Joint 3D proposal generation and object detection from view aggregation. In Proceedings of the IEEE\/RSJ International Conference on Intelligent Robots and Systems (IROS\u201918). IEEE, 1--8."},{"key":"e_1_2_1_73_1","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 3559--3568","author":"Kundu Abhijit","unstructured":"Abhijit Kundu , Yin Li , and James M. Rehg . 2018. 3D-RCNN: Instance-level 3D object reconstruction via render-and-compare . In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 3559--3568 . Abhijit Kundu, Yin Li, and James M. Rehg. 2018. 3D-RCNN: Instance-level 3D object reconstruction via render-and-compare. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 3559--3568."},{"key":"e_1_2_1_74_1","doi-asserted-by":"publisher","DOI":"10.1049\/iet-cvi.2010.0040"},{"key":"e_1_2_1_75_1","unstructured":"Fredrik Larsson Michael Felsberg and Per-Erik Forssen. 2019. Swedish Traffic Signs Dataset. Retrieved from http:\/\/www.cvl.isy.liu.se\/research\/datasets\/traffic-signs-dataset\/.  Fredrik Larsson Michael Felsberg and Per-Erik Forssen. 2019. Swedish Traffic Signs Dataset. Retrieved from http:\/\/www.cvl.isy.liu.se\/research\/datasets\/traffic-signs-dataset\/."},{"key":"e_1_2_1_76_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01264-9_45"},{"key":"e_1_2_1_77_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2006.68"},{"key":"e_1_2_1_78_1","unstructured":"Yann LeCun et\u00a0al. 2015. LeNet-5 convolutional neural networks. Retrieved from http:\/\/yann. lecun. com\/exdb\/lenet.  Yann LeCun et\u00a0al. 2015. LeNet-5 convolutional neural networks. Retrieved from http:\/\/yann. lecun. com\/exdb\/lenet."},{"key":"e_1_2_1_79_1","volume-title":"Burges","author":"LeCun Yann","year":"1998","unstructured":"Yann LeCun , Corinna Cortes , and Christopher J. C . Burges . 1998 . The MNIST database of handwritten digits. Retrieved from https:\/\/http:\/\/yann.lecun.com\/exdb\/mnist\/. Yann LeCun, Corinna Cortes, and Christopher J. C. Burges. 1998. The MNIST database of handwritten digits. Retrieved from https:\/\/http:\/\/yann.lecun.com\/exdb\/mnist\/."},{"key":"e_1_2_1_80_1","doi-asserted-by":"publisher","DOI":"10.1109\/IROS.2017.8205955"},{"key":"e_1_2_1_81_1","unstructured":"Bo Li Tianlei Zhang and Tian Xia. 2016. Vehicle detection from 3D lidar using fully convolutional network. Retrieved from https:\/\/Arxiv:1608.07916.  Bo Li Tianlei Zhang and Tian Xia. 2016. Vehicle detection from 3D lidar using fully convolutional network. Retrieved from https:\/\/Arxiv:1608.07916."},{"key":"e_1_2_1_82_1","doi-asserted-by":"publisher","DOI":"10.1109\/IJCNN.2018.8489623"},{"key":"e_1_2_1_83_1","doi-asserted-by":"publisher","DOI":"10.1109\/TMM.2017.2759508"},{"key":"e_1_2_1_84_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00783"},{"key":"e_1_2_1_85_1","unstructured":"Qingpeng Li Lichao Mou Qizhi Xu Yun Zhang and Xiao Xiang Zhu. 2018. R3-Net: A deep network for multi-oriented vehicle detection in aerial images and videos. Retrieved from https:\/\/Arxiv:1808.05560.  Qingpeng Li Lichao Mou Qizhi Xu Yun Zhang and Xiao Xiang Zhu. 2018. R 3 -Net: A deep network for multi-oriented vehicle detection in aerial images and videos. Retrieved from https:\/\/Arxiv:1808.05560."},{"key":"e_1_2_1_86_1","volume-title":"Line-CNN: End-to-end traffic line detection with line proposal unit","author":"Li Xiang","year":"2019","unstructured":"Xiang Li , Jun Li , Xiaolin Hu , and Jian Yang . 2019. Line-CNN: End-to-end traffic line detection with line proposal unit . IEEE Trans. Intell. Transport. Syst . ( 2019 ). Xiang Li, Jun Li, Xiaolin Hu, and Jian Yang. 2019. Line-CNN: End-to-end traffic line detection with line proposal unit. IEEE Trans. Intell. Transport. Syst. (2019)."},{"key":"e_1_2_1_87_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01270-0_39"},{"key":"e_1_2_1_88_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.106"},{"key":"e_1_2_1_89_1","volume-title":"Proceedings of the European Conference on Computer Vision. Springer, 740--755","author":"Lin Tsung-Yi","unstructured":"Tsung-Yi Lin , Michael Maire , Serge Belongie , James Hays , Pietro Perona , Deva Ramanan , Piotr Doll\u00e1r , and C. Lawrence Zitnick . 2014. Microsoft coco: Common objects in context . In Proceedings of the European Conference on Computer Vision. Springer, 740--755 . Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Doll\u00e1r, and C. Lawrence Zitnick. 2014. Microsoft coco: Common objects in context. In Proceedings of the European Conference on Computer Vision. Springer, 740--755."},{"key":"e_1_2_1_90_1","unstructured":"Li Liu Wanli Ouyang Xiaogang Wang Paul Fieguth Jie Chen Xinwang Liu and Matti Pietik\u00e4inen. 2018. Deep learning for generic object detection: A survey. Retrieved from Arxiv:1809.02165.  Li Liu Wanli Ouyang Xiaogang Wang Paul Fieguth Jie Chen Xinwang Liu and Matti Pietik\u00e4inen. 2018. Deep learning for generic object detection: A survey. Retrieved from Arxiv:1809.02165."},{"key":"e_1_2_1_91_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-019-01247-4"},{"key":"e_1_2_1_92_1","volume-title":"Berg","author":"Liu Wei","year":"2016","unstructured":"Wei Liu , Dragomir Anguelov , Dumitru Erhan , Christian Szegedy , Scott Reed , Cheng-Yang Fu , and Alexander C . Berg . 2016 . SSD : Single shot multibox detector. In Proceedings of the European Conference on Computer Vision. Springer , 21--37. Wei Liu, Dragomir Anguelov, Dumitru Erhan, Christian Szegedy, Scott Reed, Cheng-Yang Fu, and Alexander C. Berg. 2016. SSD: Single shot multibox detector. In Proceedings of the European Conference on Computer Vision. Springer, 21--37."},{"key":"e_1_2_1_93_1","doi-asserted-by":"publisher","DOI":"10.1109\/ITSC.2013.6728559"},{"key":"e_1_2_1_94_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2015.7298965"},{"key":"e_1_2_1_95_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.imavis.2004.02.006"},{"key":"e_1_2_1_96_1","doi-asserted-by":"publisher","DOI":"10.1109\/IRI.2017.57"},{"key":"e_1_2_1_97_1","volume-title":"An embedded computer-vision system for multi-object detection in traffic surveillance","author":"Mhalla Ala","year":"2018","unstructured":"Ala Mhalla , Thierry Chateau , Sami Gazzah , and Najoua Essoukri Ben Amara . 2018. An embedded computer-vision system for multi-object detection in traffic surveillance . IEEE Trans. Intell. Transport. Syst . ( 2018 ). Ala Mhalla, Thierry Chateau, Sami Gazzah, and Najoua Essoukri Ben Amara. 2018. An embedded computer-vision system for multi-object detection in traffic surveillance. IEEE Trans. Intell. Transport. Syst. (2018)."},{"key":"e_1_2_1_98_1","doi-asserted-by":"publisher","DOI":"10.1109\/CCDC.2018.8408156"},{"key":"e_1_2_1_99_1","unstructured":"Hans Moravec. 1988. Moravec\u2019s paradox. Retrieved from https:\/\/en.wikipedia.org\/wiki\/Moravec%27s_paradox#CITEREFMoravec1988.  Hans Moravec. 1988. Moravec\u2019s paradox. Retrieved from https:\/\/en.wikipedia.org\/wiki\/Moravec%27s_paradox#CITEREFMoravec1988."},{"key":"e_1_2_1_100_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.597"},{"key":"e_1_2_1_101_1","doi-asserted-by":"publisher","DOI":"10.1109\/ITSC.2018.8569683"},{"key":"e_1_2_1_102_1","unstructured":"The Chinese University of Hong Kong Multimedia Laboratory. 2018. CULane Dataset. Retrieved from https:\/\/xingangpan.github.io\/projects\/CULane.html.  The Chinese University of Hong Kong Multimedia Laboratory. 2018. CULane Dataset. Retrieved from https:\/\/xingangpan.github.io\/projects\/CULane.html."},{"key":"e_1_2_1_103_1","doi-asserted-by":"publisher","DOI":"10.1093\/nar\/gkg509"},{"key":"e_1_2_1_104_1","unstructured":"Karlsruhe Institute of Technology and Toyota Technological Institute at Chicago. 2012. KITTI vision benchmark suite. Retrieved from http:\/\/www.cvlibs.net\/datasets\/kitti\/eval_object.php.  Karlsruhe Institute of Technology and Toyota Technological Institute at Chicago. 2012. KITTI vision benchmark suite. Retrieved from http:\/\/www.cvlibs.net\/datasets\/kitti\/eval_object.php."},{"key":"e_1_2_1_105_1","doi-asserted-by":"publisher","DOI":"10.1109\/IROS.2016.7759717"},{"key":"e_1_2_1_106_1","volume-title":"Proceedings of the 32nd AAAI Conference on Artificial Intelligence.","author":"Pan Xingang","year":"2018","unstructured":"Xingang Pan , Jianping Shi , Ping Luo , Xiaogang Wang , and Xiaoou Tang . 2018 . Spatial as deep: Spatial cnn for traffic scene understanding . In Proceedings of the 32nd AAAI Conference on Artificial Intelligence. Xingang Pan, Jianping Shi, Ping Luo, Xiaogang Wang, and Xiaoou Tang. 2018. Spatial as deep: Spatial cnn for traffic scene understanding. In Proceedings of the 32nd AAAI Conference on Artificial Intelligence."},{"key":"e_1_2_1_107_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.sigpro.2010.08.010"},{"key":"e_1_2_1_108_1","volume-title":"Patil and Subrahmanyam Murala","author":"Prashant","year":"2018","unstructured":"Prashant W. Patil and Subrahmanyam Murala . 2018 . MSFgNet: A novel compact end-to-end deep network for moving object detection. IEEE Trans. Intell. Transport. Syst . (2018). Prashant W. Patil and Subrahmanyam Murala. 2018. MSFgNet: A novel compact end-to-end deep network for moving object detection. IEEE Trans. Intell. Transport. Syst. (2018)."},{"key":"e_1_2_1_109_1","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 918--927","author":"Qi Charles R.","unstructured":"Charles R. Qi , Wei Liu , Chenxia Wu , Hao Su , and Leonidas J. Guibas . 2018. Frustum pointnets for 3D object detection from RGB-D data . In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 918--927 . Charles R. Qi, Wei Liu, Chenxia Wu, Hao Su, and Leonidas J. Guibas. 2018. Frustum pointnets for 3D object detection from RGB-D data. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 918--927."},{"key":"e_1_2_1_110_1","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 652--660","author":"Qi Charles R.","unstructured":"Charles R. Qi , Hao Su , Kaichun Mo , and Leonidas J. Guibas . 2017. Pointnet: Deep learning on point sets for 3D classification and segmentation . In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 652--660 . Charles R. Qi, Hao Su, Kaichun Mo, and Leonidas J. Guibas. 2017. Pointnet: Deep learning on point sets for 3D classification and segmentation. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 652--660."},{"key":"e_1_2_1_111_1","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 5648--5656","author":"Qi Charles R.","unstructured":"Charles R. Qi , Hao Su , Matthias Nie\u00dfner , Angela Dai , Mengyuan Yan , and Leonidas J. Guibas . 2016. Volumetric and multi-view cnns for object classification on 3D data . In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 5648--5656 . Charles R. Qi, Hao Su, Matthias Nie\u00dfner, Angela Dai, Mengyuan Yan, and Leonidas J. Guibas. 2016. Volumetric and multi-view cnns for object classification on 3D data. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 5648--5656."},{"key":"e_1_2_1_112_1","volume-title":"Guibas","author":"Qi Charles Ruizhongtai","year":"2017","unstructured":"Charles Ruizhongtai Qi , Li Yi , Hao Su , and Leonidas J . Guibas . 2017 . Pointnet++: Deep hierarchical feature learning on point sets in a metric space. In Advances in Neural Information Processing Systems. MIT Press , 5099--5108. Charles Ruizhongtai Qi, Li Yi, Hao Su, and Leonidas J. Guibas. 2017. Pointnet++: Deep hierarchical feature learning on point sets in a metric space. In Advances in Neural Information Processing Systems. MIT Press, 5099--5108."},{"key":"e_1_2_1_113_1","doi-asserted-by":"publisher","DOI":"10.1109\/FSKD.2016.7603233"},{"key":"e_1_2_1_114_1","doi-asserted-by":"publisher","DOI":"10.1109\/CISP-BMEI.2018.8633119"},{"key":"e_1_2_1_115_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.91"},{"key":"e_1_2_1_116_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.690"},{"key":"e_1_2_1_117_1","unstructured":"Joseph Redmon and Ali Farhadi. 2018. Yolov3: An incremental improvement. Retrieved from https:\/\/Arxiv:1804.02767.  Joseph Redmon and Ali Farhadi. 2018. Yolov3: An incremental improvement. Retrieved from https:\/\/Arxiv:1804.02767."},{"key":"e_1_2_1_118_1","volume-title":"Advances in Neural Information Processing Systems","author":"Ren Shaoqing","unstructured":"Shaoqing Ren , Kaiming He , Ross Girshick , and Jian Sun . 2015. Faster R-CNN: Towards real-time object detection with region proposal networks . In Advances in Neural Information Processing Systems . MIT Press , 91--99. Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun. 2015. Faster R-CNN: Towards real-time object detection with region proposal networks. In Advances in Neural Information Processing Systems. MIT Press, 91--99."},{"key":"e_1_2_1_119_1","unstructured":"Frank Rosenblatt. 1957. The Perceptron a Perceiving and Recognizing Automaton Project Para. Cornell Aeronautical Laboratory.  Frank Rosenblatt. 1957. The Perceptron a Perceiving and Recognizing Automaton Project Para. Cornell Aeronautical Laboratory."},{"key":"e_1_2_1_120_1","doi-asserted-by":"publisher","DOI":"10.1109\/ITSC.2017.8317599"},{"key":"e_1_2_1_121_1","volume-title":"Smola","author":"Scholkopf Bernhard","year":"2001","unstructured":"Bernhard Scholkopf and Alexander J . Smola . 2001 . Learning with Kernels : Support Vector Machines, Regularization, Optimization, and Beyond. MIT Press . Bernhard Scholkopf and Alexander J. Smola. 2001. Learning with Kernels: Support Vector Machines, Regularization, Optimization, and Beyond. MIT Press."},{"key":"e_1_2_1_122_1","doi-asserted-by":"publisher","DOI":"10.1109\/IVS.2019.8813895"},{"key":"e_1_2_1_123_1","unstructured":"Karen Simonyan and Andrew Zisserman. 2014. Very deep convolutional networks for large-scale image recognition. Retrieved from https:\/\/Arxiv:1409.1556.  Karen Simonyan and Andrew Zisserman. 2014. Very deep convolutional networks for large-scale image recognition. Retrieved from https:\/\/Arxiv:1409.1556."},{"key":"e_1_2_1_124_1","doi-asserted-by":"publisher","DOI":"10.1109\/TITS.2010.2040177"},{"key":"e_1_2_1_125_1","doi-asserted-by":"publisher","DOI":"10.1109\/TITS.2013.2266661"},{"key":"e_1_2_1_126_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.94"},{"key":"e_1_2_1_127_1","unstructured":"Johannes Stallkamp Marc Schlipsing Jan Salmen and Christian Igel. 2012. German Traffic Sign Recognition Benchmark. Retrieved from http:\/\/benchmark.ini.rub.de\/?section&equals;gtsrb&subsection&equals;news.  Johannes Stallkamp Marc Schlipsing Jan Salmen and Christian Igel. 2012. German Traffic Sign Recognition Benchmark. Retrieved from http:\/\/benchmark.ini.rub.de\/?section&equals;gtsrb&subsection&equals;news."},{"key":"e_1_2_1_128_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.neunet.2012.02.016"},{"key":"e_1_2_1_129_1","volume-title":"Intelligent Computing: Image Processing Based Applications","author":"Sultana Farhana","unstructured":"Farhana Sultana , Abu Sufian , and Paramartha Dutta . 2020. A review of object detection models based on convolutional neural network . In Intelligent Computing: Image Processing Based Applications . Springer , 1--16. Farhana Sultana, Abu Sufian, and Paramartha Dutta. 2020. A review of object detection models based on convolutional neural network. In Intelligent Computing: Image Processing Based Applications. Springer, 1--16."},{"key":"e_1_2_1_130_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2006.104"},{"key":"e_1_2_1_131_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2015.7298594"},{"key":"e_1_2_1_132_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.308"},{"key":"e_1_2_1_133_1","doi-asserted-by":"publisher","DOI":"10.3390\/s17020336"},{"key":"e_1_2_1_134_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCSNT.2017.8343709"},{"key":"e_1_2_1_135_1","volume-title":"FCOS: Fully convolutional one-stage object detection.","author":"Tian Zhi","year":"2019","unstructured":"Zhi Tian , Chunhua Shen , Hao Chen , and Tong He . 2019 . FCOS: Fully convolutional one-stage object detection. Retrieved from https:\/\/Arxiv:1904.01355. Zhi Tian, Chunhua Shen, Hao Chen, and Tong He. 2019. FCOS: Fully convolutional one-stage object detection. Retrieved from https:\/\/Arxiv:1904.01355."},{"key":"e_1_2_1_136_1","unstructured":"Tusimple. 2017. Tusimple Benchmark. Retrieved from http:\/\/benchmark.tusimple.ai\/#\/.  Tusimple. 2017. Tusimple Benchmark. Retrieved from http:\/\/benchmark.tusimple.ai\/#\/."},{"key":"e_1_2_1_137_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-013-0620-5"},{"key":"e_1_2_1_138_1","doi-asserted-by":"publisher","DOI":"10.1109\/AVSS.2018.8639489"},{"key":"e_1_2_1_139_1","doi-asserted-by":"publisher","DOI":"10.15199\/48.2015.12.08"},{"key":"e_1_2_1_140_1","unstructured":"Shiyao Wang Hongchao Lu Pavel Dmitriev and Zhidong Deng. 2018. Fast object detection in compressed video. Retrieved from http:\/\/arxiv.org\/abs\/1811.11057.  Shiyao Wang Hongchao Lu Pavel Dmitriev and Zhidong Deng. 2018. Fast object detection in compressed video. Retrieved from http:\/\/arxiv.org\/abs\/1811.11057."},{"key":"e_1_2_1_141_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01261-8_33"},{"key":"e_1_2_1_142_1","unstructured":"Xiaogang Wang Xiaoxu Ma and W. Eric L. Grimson. 2019. MIT Traffic Data Set. Retrieved from http:\/\/www.ee.cuhk.edu.hk\/ xgwang\/MITtraffic.html.  Xiaogang Wang Xiaoxu Ma and W. Eric L. Grimson. 2019. MIT Traffic Data Set. Retrieved from http:\/\/www.ee.cuhk.edu.hk\/ xgwang\/MITtraffic.html."},{"key":"e_1_2_1_143_1","volume-title":"Proceedings of the 21st International Conference on Intelligent Transportation Systems (ITSC\u201918)","author":"Weber Michael","unstructured":"Michael Weber , Matthias Huber , and J. Marius Z\u00f6llner . 2018. HDTLR: A CNN-based hierarchical detector for traffic lights . In Proceedings of the 21st International Conference on Intelligent Transportation Systems (ITSC\u201918) . IEEE, 255--260. Michael Weber, Matthias Huber, and J. Marius Z\u00f6llner. 2018. HDTLR: A CNN-based hierarchical detector for traffic lights. In Proceedings of the 21st International Conference on Intelligent Transportation Systems (ITSC\u201918). IEEE, 255--260."},{"key":"e_1_2_1_144_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPRW.2017.60"},{"key":"e_1_2_1_145_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00631"},{"key":"e_1_2_1_146_1","doi-asserted-by":"publisher","DOI":"10.5555\/3196158.3196225"},{"key":"e_1_2_1_147_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00249"},{"key":"e_1_2_1_148_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00033"},{"key":"e_1_2_1_149_1","volume-title":"Car detection from low-altitude UAV imagery with the faster R-CNN. J. Adv. Transport. 2017","author":"Xu Yongzheng","year":"2017","unstructured":"Yongzheng Xu , Guizhen Yu , Yunpeng Wang , Xinkai Wu , and Yalong Ma. 2017. Car detection from low-altitude UAV imagery with the faster R-CNN. J. Adv. Transport. 2017 ( 2017 ). Yongzheng Xu, Guizhen Yu, Yunpeng Wang, Xinkai Wu, and Yalong Ma. 2017. Car detection from low-altitude UAV imagery with the faster R-CNN. J. Adv. Transport. 2017 (2017)."},{"key":"e_1_2_1_150_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00798"},{"key":"e_1_2_1_151_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.comnet.2018.02.026"},{"key":"e_1_2_1_152_1","volume-title":"Real-Time Image and Video Processing","author":"Yang Wei","year":"2018","unstructured":"Wei Yang , Ji Zhang , Hongyuan Wang , and Zhongbao Zhang . 2018. A vehicle real-time detection algorithm based on YOLOv2 framework . In Real-Time Image and Video Processing 2018 , Vol. 10670 . International Society for Optics and Photonics , 106700N. Wei Yang, Ji Zhang, Hongyuan Wang, and Zhongbao Zhang. 2018. A vehicle real-time detection algorithm based on YOLOv2 framework. In Real-Time Image and Video Processing 2018, Vol. 10670. International Society for Optics and Photonics, 106700N."},{"key":"e_1_2_1_153_1","doi-asserted-by":"publisher","DOI":"10.1109\/CIS.2017.00099"},{"key":"e_1_2_1_154_1","doi-asserted-by":"publisher","DOI":"10.3390\/a10040127"},{"key":"e_1_2_1_155_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.474"},{"key":"e_1_2_1_156_1","unstructured":"Shanshan Zhang Rodrigo Benenson and Bernt Schiele. 2019. Citypersons Data Set. Retrieved from https:\/\/bitbucket.org\/shanshanzhang\/citypersons.  Shanshan Zhang Rodrigo Benenson and Bernt Schiele. 2019. Citypersons Data Set. Retrieved from https:\/\/bitbucket.org\/shanshanzhang\/citypersons."},{"key":"e_1_2_1_157_1","volume-title":"Proceedings of the European Conference on Computer Vision (ECCV\u201918)","author":"Zhang Shifeng","unstructured":"Shifeng Zhang , Longyin Wen , Xiao Bian , Zhen Lei , and Stan Z. Li . 2018. Occlusion-aware R-CNN: Detecting pedestrians in a crowd . In Proceedings of the European Conference on Computer Vision (ECCV\u201918) . 637--653. Shifeng Zhang, Longyin Wen, Xiao Bian, Zhen Lei, and Stan Z. Li. 2018. Occlusion-aware R-CNN: Detecting pedestrians in a crowd. In Proceedings of the European Conference on Computer Vision (ECCV\u201918). 637--653."},{"key":"e_1_2_1_158_1","doi-asserted-by":"publisher","DOI":"10.1145\/2522968.2522978"},{"key":"e_1_2_1_159_1","doi-asserted-by":"publisher","DOI":"10.1109\/VTCFall.2016.7880852"},{"key":"e_1_2_1_160_1","doi-asserted-by":"publisher","DOI":"10.1109\/TNNLS.2018.2876865"},{"key":"e_1_2_1_161_1","doi-asserted-by":"publisher","DOI":"10.23919\/ChiCC.2017.8029130"},{"key":"e_1_2_1_162_1","doi-asserted-by":"crossref","unstructured":"Xingyi Zhou Jiacheng Zhuo and Philipp Kr\u00e4henb\u00fchl. 2019. Bottom-up object detection by grouping extreme and center points. Retrieved from https:\/\/Arxiv:1901.08043.  Xingyi Zhou Jiacheng Zhuo and Philipp Kr\u00e4henb\u00fchl. 2019. Bottom-up object detection by grouping extreme and center points. Retrieved from https:\/\/Arxiv:1901.08043.","DOI":"10.1109\/CVPR.2019.00094"},{"key":"e_1_2_1_163_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00472"},{"key":"e_1_2_1_164_1","unstructured":"Chenchen Zhu Yihui He and Marios Savvides. 2019. Feature selective anchor-free module for single-shot object detection. Retrieved from https:\/\/Arxiv:1903.00621.  Chenchen Zhu Yihui He and Marios Savvides. 2019. Feature selective anchor-free module for single-shot object detection. Retrieved from https:\/\/Arxiv:1903.00621."},{"key":"e_1_2_1_165_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00753"},{"key":"e_1_2_1_166_1","unstructured":"Xizhou Zhu Jifeng Dai Xingchi Zhu Yichen Wei and Lu Yuan. 2018. Towards high performance video object detection for mobiles. Retrieved from http:\/\/arxiv.org\/abs\/1804.05830.  Xizhou Zhu Jifeng Dai Xingchi Zhu Yichen Wei and Lu Yuan. 2018. Towards high performance video object detection for mobiles. Retrieved from http:\/\/arxiv.org\/abs\/1804.05830."},{"key":"e_1_2_1_167_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.52"},{"key":"e_1_2_1_168_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.441"},{"key":"e_1_2_1_169_1","doi-asserted-by":"publisher","DOI":"10.1109\/TITS.2017.2768827"},{"key":"e_1_2_1_170_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.232"},{"key":"e_1_2_1_171_1","unstructured":"Zhe Zhu Dun Liang Songhai Zhang Xiaolei Huang Baoli Li and Shimin Hu. 2016. tsinghua-tencent 100k dataset. Retrieved from https:\/\/cg.cs.tsinghua.edu.cn\/traffic-sign\/.  Zhe Zhu Dun Liang Songhai Zhang Xiaolei Huang Baoli Li and Shimin Hu. 2016. tsinghua-tencent 100k dataset. Retrieved from https:\/\/cg.cs.tsinghua.edu.cn\/traffic-sign\/."},{"key":"e_1_2_1_172_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDCSW.2017.34"}],"container-title":["ACM Computing Surveys"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3434398","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3434398","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T17:31:48Z","timestamp":1750181508000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3434398"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2021,3,5]]},"references-count":171,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2022,3,31]]}},"alternative-id":["10.1145\/3434398"],"URL":"https:\/\/doi.org\/10.1145\/3434398","relation":{},"ISSN":["0360-0300","1557-7341"],"issn-type":[{"value":"0360-0300","type":"print"},{"value":"1557-7341","type":"electronic"}],"subject":[],"published":{"date-parts":[[2021,3,5]]},"assertion":[{"value":"2019-09-01","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2020-11-01","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2021-03-05","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}