{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,29]],"date-time":"2026-04-29T12:17:03Z","timestamp":1777465023300,"version":"3.51.4"},"reference-count":39,"publisher":"Springer Science and Business Media LLC","issue":"1","license":[{"start":{"date-parts":[[2022,2,2]],"date-time":"2022-02-02T00:00:00Z","timestamp":1643760000000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2022,2,2]],"date-time":"2022-02-02T00:00:00Z","timestamp":1643760000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"DOI":"10.13039\/501100004663","name":"Ministry of Science and Technology, Taiwan","doi-asserted-by":"publisher","award":["MOST 109-2218-E-006-032"],"award-info":[{"award-number":["MOST 109-2218-E-006-032"]}],"id":[{"id":"10.13039\/501100004663","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100005144","name":"Qualcomm","doi-asserted-by":"publisher","award":["SOW#NAT-435536"],"award-info":[{"award-number":["SOW#NAT-435536"]}],"id":[{"id":"10.13039\/100005144","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["EURASIP J. Adv. Signal Process."],"published-print":{"date-parts":[[2022,12]]},"abstract":"<jats:title>Abstract<\/jats:title>\n                  <jats:p>The vision-based smart driving technologies for road safety are the popular research topics in computer vision. The precise moving object detection with continuously tracking capability is one of the most important vision-based technologies nowadays. In this paper, we propose an improved object detection system, which combines a typical object detector and long short-term memory (LSTM) modules, to further improve the detection performance for smart driving. First, starting from a selected object detector, we combine all vehicle classes and bypassing low-level features to improve its detection performance. After the spatial association of the detected objects, the outputs of the improved object detector are then fed into the proposed double-layer LSTM (dLSTM) modules to successfully improve the detection performance of the vehicles in various conditions, including the newly-appeared, the detected and the gradually-disappearing vehicles. With stage-by-stage evaluations, the experimental results show that the proposed vehicle detection system with dLSTM modules can precisely detect the vehicles without increasing computations.<\/jats:p>","DOI":"10.1186\/s13634-022-00839-6","type":"journal-article","created":{"date-parts":[[2022,2,2]],"date-time":"2022-02-02T15:53:52Z","timestamp":1643817232000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":11,"title":["Improved vehicle detection systems with double-layer LSTM modules"],"prefix":"10.1186","volume":"2022","author":[{"given":"Wei-Jong","family":"Yang","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Wan-Ju","family":"Liow","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Shao-Fu","family":"Chen","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-3024-5634","authenticated-orcid":false,"given":"Jar-Ferr","family":"Yang","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Pau-Choo","family":"Chung","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Songan","family":"Mao","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2022,2,2]]},"reference":[{"key":"839_CR1","doi-asserted-by":"crossref","unstructured":"Z. Cao, G. Hidalgo, T. Sion, S.-E. Wei, and Y. Sheikh, OpenPose: realtime multi-person 2D pose estimation using part affinity fields (2018). arXiv:1812.08008.","DOI":"10.1109\/CVPR.2017.143"},{"key":"839_CR2","doi-asserted-by":"crossref","unstructured":"L. Pishchulin, E. Insafutdinov, S. Tang, B. Andres, M. Andriluka, P. V. Gehler, and B. Schiele, Deepcut: Joint subset partition and labeling for multi person pose estimation, in Proceedings of IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp. 4929\u20134937 (2016).","DOI":"10.1109\/CVPR.2016.533"},{"issue":"10","key":"839_CR3","doi-asserted-by":"publisher","first-page":"1499","DOI":"10.1109\/LSP.2016.2603342","volume":"23","author":"K Zhang","year":"2016","unstructured":"K. Zhang, Z. Zhang, Z. Li, Y. Qiao, Joint face detection and alignment using multitask cascaded convolutional networks. IEEE Signal Process. Lett. 23(10), 1499\u20131503 (2016)","journal-title":"IEEE Signal Process. Lett."},{"key":"839_CR4","doi-asserted-by":"crossref","unstructured":"Y. Taigman, M. Yang, M. A. Ranzato, and L. Wolf, Deepface: closing the gap to human-level performance in face verification, in Proceedings of IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp. 1701\u20131708 (2014).","DOI":"10.1109\/CVPR.2014.220"},{"key":"839_CR5","doi-asserted-by":"crossref","unstructured":"J. Long, E. Shelhamer, and T. Darrell, Fully convolutional networks for semantic segmentation, in Proceedings of IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp. 3431\u20133440 (2015).","DOI":"10.1109\/CVPR.2015.7298965"},{"key":"839_CR6","doi-asserted-by":"crossref","unstructured":"K. He, G. Gkioxari, P. Doll\u00e1r, and R. Girshick, Mask r-cnn, in Proceedings of IEEE Conference on Computer Vision (ICCV), pp. 2961\u20132969 (2017).","DOI":"10.1109\/ICCV.2017.322"},{"issue":"12","key":"839_CR7","doi-asserted-by":"publisher","first-page":"2481","DOI":"10.1109\/TPAMI.2016.2644615","volume":"39","author":"V Badrinarayanan","year":"2017","unstructured":"V. Badrinarayanan, A. Kendall, R. Cipolla, Segnet: a deep convolutional encoder\u2013decoder architecture for image segmentation. IEEE Trans. Pattern Anal. Mach. Intell. 39(12), 2481\u20132495 (2017)","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"839_CR8","doi-asserted-by":"crossref","unstructured":"N. Dalal, and B. Triggs, Histograms of oriented gradients for human detection, in Proceedings of IEEE Conference on Computer Vision and Pattern Recognition, pp. 886\u2013893 (2015).","DOI":"10.1109\/CVPR.2005.177"},{"key":"839_CR9","doi-asserted-by":"crossref","unstructured":"D. G. Lowe, Object recognition from local scale-invariant features, in Proceedings of IEEE Conference on Computer Vision (ICCV), pp. 1150\u20131157 (1999).","DOI":"10.1109\/ICCV.1999.790410"},{"issue":"9","key":"839_CR10","doi-asserted-by":"publisher","first-page":"1627","DOI":"10.1109\/TPAMI.2009.167","volume":"32","author":"PF Felzenszwalb","year":"2009","unstructured":"P.F. Felzenszwalb, R.B. Girshick, D. McAllester, D. Ramanan, Object detection with discriminatively trained part-based models. IEEE Trans. Pattern Anal. Mach. Intell. 32(9), 1627\u20131645 (2009)","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"839_CR11","first-page":"511","volume":"1","author":"P Viola","year":"2001","unstructured":"P. Viola, M. Jones, Rapid object detection using a boosted cascade of simple features. Proc. IEEE Conf. Comput. Vis. Pattern Recognit. 1, 511\u2013518 (2001)","journal-title":"Proc. IEEE Conf. Comput. Vis. Pattern Recognit."},{"issue":"1","key":"839_CR12","doi-asserted-by":"publisher","first-page":"119","DOI":"10.1006\/jcss.1997.1504","volume":"55","author":"Y Freund","year":"1997","unstructured":"Y. Freund, R.E. Schapire, A decision-theoretic generalization of on-line learning and an application to boosting. J. Comput. Syst. Sci. 55(1), 119\u2013139 (1997)","journal-title":"J. Comput. Syst. Sci."},{"key":"839_CR13","unstructured":"C.-W. Hsu, C.-C. Chang, and C.-J. Lin, A practical guide to support vector classification. Technical Guide, Department of Computer Science, National Taiwan University, Taipei 106, Taiwan. http:\/\/www.csie.ntu.edu.tw\/~cjlin2010."},{"issue":"2","key":"839_CR14","doi-asserted-by":"publisher","first-page":"303","DOI":"10.1007\/s11263-009-0275-4","volume":"88","author":"M Everingham","year":"2010","unstructured":"M. Everingham, L. Van Gool, C.K. Williams, J. Winn, A. Zisserman, The pascal visual object classes (voc) challenge. Int. J. Comput. Vis. 88(2), 303\u2013338 (2010)","journal-title":"Int. J. Comput. Vis."},{"key":"839_CR15","unstructured":"A. Krizhevsky, I. Sutskever, and G. E. Hinton, Imagenet classification with deep convolutional neural networks, in Proceedings of Neural Information Processing Systems Conf. (NIPS), pp. 1097\u20131105 (2012)."},{"key":"839_CR16","doi-asserted-by":"crossref","unstructured":"M. D. Zeiler, and R. Fergus, Visualizing and understanding convolutional networks, in Proceedings of European Conference on Computer Vision (ECCV), pp. 818\u2013833 (2014).","DOI":"10.1007\/978-3-319-10590-1_53"},{"key":"839_CR17","doi-asserted-by":"crossref","unstructured":"C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich, Going deeper with convolutions, in Proceedings of IEEE Conference on Computer Vision and Pattern Recognition, pp. 1\u20139 (2015).","DOI":"10.1109\/CVPR.2015.7298594"},{"key":"839_CR18","doi-asserted-by":"crossref","unstructured":"K. He, X. Zhang, S. Ren, and J. Sun, Deep residual learning for image recognition, in Proceeedings of IEEE Conference on Computer Vision and Pattern Recognition, pp. 770\u2013778 (2016).","DOI":"10.1109\/CVPR.2016.90"},{"key":"839_CR19","doi-asserted-by":"crossref","unstructured":"J. Hu, L. Shen, and G. Sun, Squeeze-and-excitation networks, in Proceedings of IEEE Conference on Computer Vision and Pattern Recognition, pp. 7132\u20137141 (2018).","DOI":"10.1109\/CVPR.2018.00745"},{"key":"839_CR20","doi-asserted-by":"crossref","unstructured":"J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, Imagenet: a large-scale hierarchical image database, in Proceedings of IEEE Conference on Computer Vision and Pattern Recognition, pp. 248\u2013255 (2009).","DOI":"10.1109\/CVPR.2009.5206848"},{"key":"839_CR21","doi-asserted-by":"crossref","unstructured":"R. Girshick, J. Donahue, T. Darrell, and J. Malik, Rich feature hierarchies for accurate object detection and semantic segmentation, in Proceedings of IEEE Conference on Computer Vision and Pattern Recognition, pp. 580\u2013587 (2014).","DOI":"10.1109\/CVPR.2014.81"},{"key":"839_CR22","doi-asserted-by":"crossref","unstructured":"R. Girshick, Fast r-cnn, in Proceedings of IEEE Conference on Computer Vision (ICCV), pp. 1440\u20131448 (2015).","DOI":"10.1109\/ICCV.2015.169"},{"key":"839_CR23","unstructured":"S. Ren, K. He, R. Girshick, and J. Sun, Faster r-cnn: towards real-time object detection with region proposal networks, in Procedings of Neural Information Processing Systems Conference (NIPS), pp. 91\u201399 (2015)."},{"key":"839_CR24","doi-asserted-by":"crossref","unstructured":"W. Liu, D. Anguelov, D. Erhan, C. Szegedy, S. Reed, C.-Y. Fu, and A. C. Berg, SSD: single shot multibox detector, in Proceedings of European Conference on Computer Vision, pp. 21\u201337 (2016).","DOI":"10.1007\/978-3-319-46448-0_2"},{"key":"839_CR25","doi-asserted-by":"crossref","unstructured":"J. Redmon, S. Divvala, R. Girshick, and A. Farhadi, You only look once: unified, real-time object detection, in Proceedings of IEEE Conference on Computer Vision and Pattern Recognition, pp. 779\u2013788 (2016).","DOI":"10.1109\/CVPR.2016.91"},{"key":"839_CR26","doi-asserted-by":"crossref","unstructured":"Q. Chu, W. Ouyang, H. Li, X. Wang, B. Liu and N. Yu. Online multi-object tracking using CNN-based single object tracker with spatial-temporal attention mechanism, in Proceedings of the IEEE International Conference on Computer Vision, pp. 4836\u20134845, (2017).","DOI":"10.1109\/ICCV.2017.518"},{"key":"839_CR27","doi-asserted-by":"crossref","unstructured":"B. Li, J. Yan, W. Wu, Z. Zhu, and X. Hu, High performance visual tracking with siamese region proposal network, in Proceedings of IEEE Conference on Computer Vision and Pattern Recognition, pp. 8971\u20138980 (2018).","DOI":"10.1109\/CVPR.2018.00935"},{"key":"839_CR28","doi-asserted-by":"crossref","unstructured":"X. Wang, R. Girshick, A. Gupta, and K. He, Non-local neural networks, in Proceedings of IEEE Conference on Computer Vision and Pattern Recognition, pp. 7794\u20137803 (2018).","DOI":"10.1109\/CVPR.2018.00813"},{"issue":"8","key":"839_CR29","doi-asserted-by":"publisher","first-page":"1735","DOI":"10.1162\/neco.1997.9.8.1735","volume":"9","author":"S Hochreiter","year":"1997","unstructured":"S. Hochreiter, J. Schmidhuber, Long short-term memory. Neural Comput. 9(8), 1735\u20131780 (1997)","journal-title":"Neural Comput."},{"key":"839_CR30","doi-asserted-by":"crossref","unstructured":"J. Redmon, and A. Farhadi, YOLO9000: better, faster, stronger, in Proceedings of IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp. 7263\u20137271 (2017).","DOI":"10.1109\/CVPR.2017.690"},{"key":"839_CR31","unstructured":"J. Redmon and A. Farhadi, YOLOv3: An incremental improvement, in Proceedings of Computer Vision and Pattern Recognition (2018). arXiv:1804.02767."},{"key":"839_CR32","unstructured":"A. Bochkovskiy, C.-Y. Wang, H.-Y. M. Liao, YOLOv4: optimal speed and accuracy of object detection, in Proceedings of Computer Vision and Pattern Recognition (2020). arXiv:2004.10934."},{"key":"839_CR33","doi-asserted-by":"crossref","unstructured":"A. Graves, A.-R. Mohamed, and G. Hinton, Speech recognition with deep recurrent neural networks, in Proceedings of IEEE International Conference on Acoustics, Speech and Signal Processing, pp. 6645\u20136649 (2013).","DOI":"10.1109\/ICASSP.2013.6638947"},{"key":"839_CR34","unstructured":"M. Hermans, and B. Schrauwen, Training and analysing deep recurrent neural networks, in Proceedings of Neural Information Processing Systems Conference, pp. 190\u2013198 (2013)."},{"key":"839_CR35","doi-asserted-by":"crossref","unstructured":"K. Lee, I. Lee and S. Lee. Propagating LSTM: 3D pose estimation based on joint interdependency, in Proceedings of European Conference on Computer Vision, pp. 119\u2013135 (2018).","DOI":"10.1007\/978-3-030-01234-2_8"},{"key":"839_CR36","doi-asserted-by":"crossref","unstructured":"S. Yun and S Kim, Recurrent YOLO and LSTM-based IR single pedestrian tracking, in Proceedings of 19th International Conference on Control, Automation and Systems, Jeju, Korea (2019).","DOI":"10.23919\/ICCAS47443.2019.8971679"},{"key":"839_CR37","doi-asserted-by":"publisher","unstructured":"G. Ning, Z. Zhang, C. Huang, Z. He, X. Ren and H. Wang, Spatially supervised recurrent convolutional neural networks for visual object tracking, in Proceedings of IEEE International Symposium on Circuits and Systems, Baltimore, MD, pp. 1\u20134 (2017). https:\/\/doi.org\/10.1109\/ISCAS.2017.8050867.","DOI":"10.1109\/ISCAS.2017.8050867"},{"key":"839_CR38","unstructured":"V. Nair, and G. E. Hinton, Rectified linear units improve restricted Boltzmann machines, in Proceedings of the 27th International Conference on Machine Learning, pp. 807\u2013814 (2010)."},{"key":"839_CR39","unstructured":"S. Ioffe, and C. Szegedy, Batch normalization: accelerating deep network training by reducing internal covariate shift (2015). arXiv:1502.03167."}],"container-title":["EURASIP Journal on Advances in Signal Processing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1186\/s13634-022-00839-6.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1186\/s13634-022-00839-6\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1186\/s13634-022-00839-6.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,9,17]],"date-time":"2024-09-17T16:11:43Z","timestamp":1726589503000},"score":1,"resource":{"primary":{"URL":"https:\/\/asp-eurasipjournals.springeropen.com\/articles\/10.1186\/s13634-022-00839-6"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,2,2]]},"references-count":39,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2022,12]]}},"alternative-id":["839"],"URL":"https:\/\/doi.org\/10.1186\/s13634-022-00839-6","relation":{"has-preprint":[{"id-type":"doi","id":"10.21203\/rs.3.rs-494794\/v1","asserted-by":"object"}]},"ISSN":["1687-6180"],"issn-type":[{"value":"1687-6180","type":"electronic"}],"subject":[],"published":{"date-parts":[[2022,2,2]]},"assertion":[{"value":"5 May 2021","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"18 January 2022","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"2 February 2022","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"In this research, we declare that the studies do not involve any human participants, human data, human tissue, and animals.","order":2,"name":"Ethics","group":{"name":"EthicsHeading","label":"Ethics approval and consent to participate"}},{"value":"In this research, we declare that the manuscript does not contain any individual person\u2019s data in any form (including individual details, images or videos).","order":3,"name":"Ethics","group":{"name":"EthicsHeading","label":"Consent for publication"}},{"value":"The authors declare that they have no competing interests.","order":4,"name":"Ethics","group":{"name":"EthicsHeading","label":"Competing interests"}}],"article-number":"7"}}