{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,2,14]],"date-time":"2026-02-14T15:07:45Z","timestamp":1771081665500,"version":"3.50.1"},"reference-count":69,"publisher":"MDPI AG","issue":"1","license":[{"start":{"date-parts":[[2025,1,17]],"date-time":"2025-01-17T00:00:00Z","timestamp":1737072000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"Key Research and Design Plan of Chinese Academy of Sciences","award":["2022YFB4400404"],"award-info":[{"award-number":["2022YFB4400404"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Information"],"abstract":"<jats:p>Deep learning significantly advances object detection. Post processes, a critical component of this process, select valid bounding boxes to represent the true targets during inference and assign boxes and labels to these objects during training to optimize the loss function. However, post processes constitute a substantial portion of the total processing time for a single image. This inefficiency primarily arises from the extensive Intersection over Union (IoU) calculations required between numerous redundant bounding boxes in post processing algorithms. To reduce these redundant IoU calculations, we introduce a classification prioritization strategy during both training and inference post processes. Additionally, post processes involve sorting operations that contribute to their inefficiency. To minimize unnecessary comparisons in Top-K sorting, we have improved the bitonic sorter by developing a hybrid bitonic algorithm. These improvements have effectively accelerated the post processing. Given the similarities between the training and inference post processes, we unify four typical post processing algorithms and design a hardware accelerator based on this framework. Our accelerator achieves at least 7.55 times the speed in inference post processing compared to that of recent accelerators. When compared to the RTX 2080 Ti system, our proposed accelerator offers at least 21.93 times the speed for the training post process and 19.89 times for the inference post process, thereby significantly enhancing the efficiency of loss function minimization.<\/jats:p>","DOI":"10.3390\/info16010063","type":"journal-article","created":{"date-parts":[[2025,1,17]],"date-time":"2025-01-17T11:24:56Z","timestamp":1737113096000},"page":"63","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":2,"title":["Object Detection Post Processing Accelerator Based on Co-Design of Hardware and Software"],"prefix":"10.3390","volume":"16","author":[{"given":"Dengtian","family":"Yang","sequence":"first","affiliation":[{"name":"Institute of Microelectronics of the Chinese Academy of Sciences, Beijing 100029, China"},{"name":"University of Chinese Academy of Sciences, Beijing 100049, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-8121-657X","authenticated-orcid":false,"given":"Lan","family":"Chen","sequence":"additional","affiliation":[{"name":"Institute of Microelectronics of the Chinese Academy of Sciences, Beijing 100029, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xiaoran","family":"Hao","sequence":"additional","affiliation":[{"name":"Institute of Microelectronics of the Chinese Academy of Sciences, Beijing 100029, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yiheng","family":"Zhang","sequence":"additional","affiliation":[{"name":"Institute of Microelectronics of the Chinese Academy of Sciences, Beijing 100029, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1968","published-online":{"date-parts":[[2025,1,17]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","unstructured":"Hu, Y., Yang, J., Chen, L., Li, K., Sima, C., Zhu, X., Chai, S., Du, S., Lin, T., and Wang, W. (2023, January 17\u201323). Planning-oriented Autonomous Driving. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Vancouver, BC, Canada.","DOI":"10.1109\/CVPR52729.2023.01712"},{"key":"ref_2","doi-asserted-by":"crossref","unstructured":"Zhou, X., Lin, Z., Shan, X., Wang, Y., Sun, D., and Yang, M.H. (2024, January 17\u201321). DrivingGaussian: Composite Gaussian Splatting for Surrounding Dynamic Autonomous Driving Scenes. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Seattle, WA, USA.","DOI":"10.1109\/CVPR52733.2024.02044"},{"key":"ref_3","doi-asserted-by":"crossref","first-page":"111402","DOI":"10.1016\/j.rse.2019.111402","article-title":"Remote sensing for agricultural applications: A meta-review","volume":"236","author":"Weiss","year":"2020","journal-title":"Remote Sens. Environ."},{"key":"ref_4","doi-asserted-by":"crossref","first-page":"100227","DOI":"10.1016\/j.atech.2023.100227","article-title":"Making technological innovations accessible to agricultural water management: Design of a low-cost wireless sensor network for drip irrigation monitoring in Tunisia","volume":"4","author":"Vandome","year":"2023","journal-title":"Smart Agric. Technol."},{"key":"ref_5","doi-asserted-by":"crossref","first-page":"106900","DOI":"10.1016\/j.gexplo.2021.106900","article-title":"Mineral prospecting from biogeochemical and geological information using hyperspectral remote sensing\u2014Feasibility and challenges","volume":"232","author":"Chakraborty","year":"2022","journal-title":"J. Geochem. Explor."},{"key":"ref_6","doi-asserted-by":"crossref","unstructured":"Wang, R., Chen, F., Wang, J., Hao, X., Chen, H., and Liu, H. (2024). Prospecting criteria for skarn-type iron deposits in the thick overburden area of Qihe-Yucheng mineral-rich area using geological and geophysical modelling. J. Appl. Geophys., 105442.","DOI":"10.1016\/j.jappgeo.2024.105442"},{"key":"ref_7","doi-asserted-by":"crossref","unstructured":"Liu, Z., Lin, Y., Cao, Y., Hu, H., Wei, Y., Zhang, Z., Lin, S., and Guo, B. (2021, January 11\u201317). Swin transformer: Hierarchical vision transformer using shifted windows. Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV), Montreal, QC, Canada.","DOI":"10.1109\/ICCV48922.2021.00986"},{"key":"ref_8","doi-asserted-by":"crossref","unstructured":"Lee, Y., Hwang, J., Lee, S., Bae, Y., and Park, J. (2019, January 16\u201320). An energy and GPU-computation efficient backbone network for real-time object detection. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPR), Long Beach, CA, USA.","DOI":"10.1109\/CVPRW.2019.00103"},{"key":"ref_9","doi-asserted-by":"crossref","unstructured":"Liu, S., Qi, L., Qin, H., Shi, J., and Jia, J. (2018, January 18\u201322). Path aggregation network for instance segmentation. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00913"},{"key":"ref_10","doi-asserted-by":"crossref","unstructured":"Lin, T.Y., Doll\u00e1r, P., Girshick, R., He, K., Hariharan, B., and Belongie, S. (2017, January 21\u201326). Feature pyramid networks for object detection. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.106"},{"key":"ref_11","doi-asserted-by":"crossref","unstructured":"Duan, K., Bai, S., Xie, L., Qi, H., Huang, Q., and Tian, Q. (2019, January 27\u201330). Centernet: Keypoint triplets for object detection. Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV), Seoul, Republic of Korea.","DOI":"10.1109\/ICCV.2019.00667"},{"key":"ref_12","first-page":"21002","article-title":"Generalized focal loss: Learning qualified and distributed bounding boxes for dense object detection","volume":"33","author":"Li","year":"2020","journal-title":"Adv. Neural Inf. Process. Syst."},{"key":"ref_13","doi-asserted-by":"crossref","unstructured":"Li, X., Wang, W., Hu, X., Li, J., Tang, J., and Yang, J. (2021, January 16\u201322). Generalized focal loss v2: Learning reliable localization quality estimation for dense object detection. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Nashville, TN, USA.","DOI":"10.1109\/CVPR46437.2021.01146"},{"key":"ref_14","doi-asserted-by":"crossref","unstructured":"Zhang, S., Chi, C., Yao, Y., Lei, Z., and Li, S.Z. (2020, January 14\u201319). Bridging the gap between anchor-based and anchor-free detection via adaptive training sample selection. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Seattle, WA, USA.","DOI":"10.1109\/CVPR42600.2020.00978"},{"key":"ref_15","doi-asserted-by":"crossref","unstructured":"Carion, N., Massa, F., Synnaeve, G., Usunier, N., Kirillov, A., and Zagoruyko, S. (2020, January 23\u201328). End-to-end object detection with transformers. Proceedings of the European Conference on Computer Vision (ECCV), Glasgow, UK.","DOI":"10.1007\/978-3-030-58452-8_13"},{"key":"ref_16","doi-asserted-by":"crossref","first-page":"7671","DOI":"10.1109\/TCSVT.2023.3277621","article-title":"Diag-IoU loss for object detection","volume":"33","author":"Zhang","year":"2023","journal-title":"IEEE Trans. Circuits Syst. Video Technol."},{"key":"ref_17","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/3570928","article-title":"FlexCNN: An end-to-end framework for composing CNN accelerators on FPGA","volume":"16","author":"Basalama","year":"2023","journal-title":"ACM Trans. Reconfigurable Technol. Syst."},{"key":"ref_18","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/3617836","article-title":"XVDPU: A High-Performance CNN Accelerator on the Versal Platform Powered by the AI Engine","volume":"17","author":"Jia","year":"2024","journal-title":"ACM Trans. Reconfigurable Technol. Syst."},{"key":"ref_19","first-page":"1","article-title":"Low-precision floating-point arithmetic for high-performance FPGA-based CNN acceleration","volume":"15","author":"Wu","year":"2021","journal-title":"ACM Trans. Reconfigurable Technol. Syst. (TRETS)"},{"key":"ref_20","doi-asserted-by":"crossref","first-page":"936","DOI":"10.1109\/TVLSI.2021.3060041","article-title":"SWM: A high-performance sparse-winograd matrix multiplication CNN accelerator","volume":"29","author":"Wu","year":"2021","journal-title":"IEEE Trans. Very Large Scale Integr. (VLSI) Syst."},{"key":"ref_21","doi-asserted-by":"crossref","first-page":"1137","DOI":"10.1109\/TPAMI.2016.2577031","article-title":"Faster R-CNN: Towards real-time object detection with region proposal networks","volume":"39","author":"Ren","year":"2016","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_22","unstructured":"(2024, January 08). NanoDet-Plus: Super Fast and High Accuracy Lightweight Anchor-Free Object Detection Model. Available online: https:\/\/github.com\/RangiLyu\/nanodet."},{"key":"ref_23","doi-asserted-by":"crossref","unstructured":"Feng, C., Zhong, Y., Gao, Y., Scott, M., and Huang, W. (2021, January 10\u201317). Tood: Task-aligned one-stage object detection. Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV), Montreal, QC, Canada.","DOI":"10.1109\/ICCV48922.2021.00349"},{"key":"ref_24","unstructured":"Lin, J., Zhu, L., Chen, W., Wang, W., Gan, C., and Han, S. (December, January 28). On-device training under 256kb memory. Proceedings of the Advances in Neural Information Processing Systems (NIPS), New Orleans, LA, USA."},{"key":"ref_25","doi-asserted-by":"crossref","first-page":"134","DOI":"10.1109\/TNSE.2021.3054583","article-title":"Resource-constrained neural architecture search on edge devices","volume":"9","author":"Lyu","year":"2021","journal-title":"IEEE Trans. Netw. Sci. Eng."},{"key":"ref_26","doi-asserted-by":"crossref","first-page":"3212","DOI":"10.1109\/TNNLS.2018.2876865","article-title":"Object detection with deep learning: A review","volume":"30","author":"Zhao","year":"2019","journal-title":"IEEE Trans. Neural Netw. Learn. Syst."},{"key":"ref_27","doi-asserted-by":"crossref","unstructured":"Tan, M., Pang, R., and Le, Q. (2020, January 14\u201319). Efficientdet: Scalable and efficient object detection. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Seattle, WA, USA.","DOI":"10.1109\/CVPR42600.2020.01079"},{"key":"ref_28","doi-asserted-by":"crossref","unstructured":"Zhao, Y., Lv, W., Xu, S., Wei, J., Wang, G., Dang, Q., Liu, Y., and Chen, J. (2024, January 17\u201321). Detrs beat yolos on real-time object detection. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Seattle, WA, USA.","DOI":"10.1109\/CVPR52733.2024.01605"},{"key":"ref_29","doi-asserted-by":"crossref","unstructured":"Liu, W., Anguelov, D., Erhan, D., Szegedy, C., Reed, S., Fu, C., and Berg, A. (2016, January 10\u201316). SSD: Single shot multibox detector. Proceedings of the 14th European Conference on Computer Vision (ECCV), Amsterdam, The Netherlands.","DOI":"10.1007\/978-3-319-46448-0_2"},{"key":"ref_30","unstructured":"Ross, T.Y., Doll\u00e1r, P., He, K., Hariharan, B., and Girshick, R. (2017, January 21\u201326). Focal loss for dense object detection. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Honolulu, HI, USA."},{"key":"ref_31","doi-asserted-by":"crossref","unstructured":"Cai, Z., and Vasconcelos, N. (2018, January 18\u201322). Cascade R-CNN: Delving into high quality object detection. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00644"},{"key":"ref_32","doi-asserted-by":"crossref","unstructured":"Law, H., and Deng, J. (2018, January 8\u201314). CornerNet: Detecting objects as paired keypoints. Proceedings of the European Conference on Computer Vision (ECCV), Munich, Germany.","DOI":"10.1007\/978-3-030-01264-9_45"},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"Zhou, X., Zhuo, J., and Krahenbuhl, P. (2019, January 16\u201320). Bottom-up object detection by grouping extreme and center points. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00094"},{"key":"ref_34","doi-asserted-by":"crossref","unstructured":"Liu, Z., Zheng, T., Xu, G., Yang, Z., Liu, H., and Cai, D. (2020, January 7\u201312). Training-time-friendly network for real-time object detection. Proceedings of the AAAI Conference on Artificial Intelligence (AAAI), New York, NY, USA.","DOI":"10.1609\/aaai.v34i07.6838"},{"key":"ref_35","doi-asserted-by":"crossref","first-page":"6780","DOI":"10.1109\/TITS.2023.3258683","article-title":"Object detection in traffic videos: A survey","volume":"24","author":"Ghahremannezhad","year":"2023","journal-title":"IEEE Trans. Intell. Transp. Syst."},{"key":"ref_36","doi-asserted-by":"crossref","unstructured":"Geng, X., Su, Y., Cao, X., Li, H., and Liu, L. (2024). YOLOFM: An improved fire and smoke object detection algorithm based on YOLOv5n. Sci. Rep., 14.","DOI":"10.1038\/s41598-024-55232-0"},{"key":"ref_37","doi-asserted-by":"crossref","unstructured":"Tian, Z., Shen, C., Chen, H., and He, T. (2019). FCOS: Fully convolutional one-stage object detection. arXiv.","DOI":"10.1109\/ICCV.2019.00972"},{"key":"ref_38","doi-asserted-by":"crossref","unstructured":"Xu, C., Wang, J., Yang, W., Yu, H., Yu, L., and Xia, G. (2022, January 23\u201327). RFLA: Gaussian receptive field based label assignment for tiny object detection. Proceedings of the European Conference on Computer Vision (ECCV), Tel-Aviv, Israel.","DOI":"10.1007\/978-3-031-20077-9_31"},{"key":"ref_39","doi-asserted-by":"crossref","unstructured":"Kim, K., and Lee, H. (2020, January 23\u201328). Probabilistic anchor assignment with IoU prediction for object detection. Proceedings of the 16th European Conference on Computer Vision (ECCV), Glasgow, UK.","DOI":"10.1007\/978-3-030-58595-2_22"},{"key":"ref_40","doi-asserted-by":"crossref","unstructured":"Ge, Z., Liu, S., Li, Z., Yoshie, O., and Sun, J. (2021, January 19\u201325). OTA: Optimal transport assignment for object detection. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Nashville, TN, USA.","DOI":"10.1109\/CVPR46437.2021.00037"},{"key":"ref_41","unstructured":"Ge, Z., Liu, S., Wang, F., Li, Z., and Sun, J. (2021). YOLOX: Exceeding YOLO series in 2021. arXiv."},{"key":"ref_42","unstructured":"Zhu, X., Su, W., Lu, L., Li, B., Wang, X., and Dai, J. (2020). Deformable DETR: Deformable transformers for end-to-end object detection. arXiv."},{"key":"ref_43","unstructured":"Pu, Y., Liang, W., Hao, Y., Yuan, Y., Yang, Y., Zhang, C., Hu, H., and Huang, G. (2023, January 10\u201316). Rank-DETR for high quality object detection. Proceedings of the 37th Conference on Neural Information Processing Systems (NeurIPS 2023), New Orleans, LA, USA."},{"key":"ref_44","doi-asserted-by":"crossref","unstructured":"Giroux, J., Bouchard, M., and Laganiere, R. (2023, January 2\u20136). T-fftradnet: Object detection with Swin vision transformers from raw ADC radar signals. Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV), Paris, France.","DOI":"10.1109\/ICCVW60793.2023.00435"},{"key":"ref_45","doi-asserted-by":"crossref","first-page":"126779","DOI":"10.1016\/j.neucom.2023.126779","article-title":"Dual Swin-transformer based mutual interactive network for RGB-D salient object detection","volume":"559","author":"Zeng","year":"2023","journal-title":"Neurocomputing"},{"key":"ref_46","doi-asserted-by":"crossref","unstructured":"Bodla, N., Singh, B., Chellappa, R., and Davis, L.S. (2017, January 22\u201329). Soft-NMS\u2013improving object detection with one line of code. Proceedings of the IEEE International Conference on Computer Vision (ICCV), Venice, Italy.","DOI":"10.1109\/ICCV.2017.593"},{"key":"ref_47","doi-asserted-by":"crossref","unstructured":"Zheng, Z., Wang, P., Liu, W., Li, J., Ye, R., and Ren, D. (2020, January 7\u201312). Distance-IoU loss: Faster and better learning for bounding box regression. Proceedings of the AAAI Conference on Artificial Intelligence (AAAI), New York, NY, USA.","DOI":"10.1609\/aaai.v34i07.6999"},{"key":"ref_48","doi-asserted-by":"crossref","first-page":"8574","DOI":"10.1109\/TCYB.2021.3095305","article-title":"Enhancing geometric factors in model learning and inference for object detection and instance segmentation","volume":"52","author":"Zheng","year":"2021","journal-title":"IEEE Trans. Cybern."},{"key":"ref_49","doi-asserted-by":"crossref","unstructured":"Rezatofighi, H., Tsoi, N., Gwak, J., Sadeghian, A., Reid, I., and Savarese, S. (2019, January 16\u201320). Generalized intersection over union: A metric and a loss for bounding box regression. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00075"},{"key":"ref_50","doi-asserted-by":"crossref","unstructured":"Neubeck, A., and Van, G.L. (2006, January 20\u201324). Efficient non-maximum suppression. Proceedings of the 18th International Conference on Pattern Recognition (ICPR), Hong Kong, China.","DOI":"10.1109\/ICPR.2006.479"},{"key":"ref_51","doi-asserted-by":"crossref","first-page":"105478","DOI":"10.1016\/j.asoc.2019.05.005","article-title":"Improved non-maximum suppression for object detection using harmony search algorithm","volume":"81","author":"Song","year":"2019","journal-title":"Appl. Soft Comput."},{"key":"ref_52","doi-asserted-by":"crossref","unstructured":"Chu, X., Zheng, A., Zhang, X., and Sun, J. (2020, January 13\u201319). Detection in crowded scenes: One proposal, multiple predictions. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Seattle, WA, USA.","DOI":"10.1109\/CVPR42600.2020.01223"},{"key":"ref_53","doi-asserted-by":"crossref","unstructured":"Liu, S., Huang, D., and Wang, Y. (2019, January 16\u201320). Adaptive NMS: Refining pedestrian detection in a crowd. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00662"},{"key":"ref_54","doi-asserted-by":"crossref","unstructured":"Gao, P., Zheng, M., Wang, X., Dai, J., and Li, H. (2021, January 10\u201317). Fast convergence of DETR with spatially modulated co-attention. Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV), Montreal, QC, Canada.","DOI":"10.1109\/ICCV48922.2021.00360"},{"key":"ref_55","doi-asserted-by":"crossref","unstructured":"Sun, P., Zhang, R., Jiang, Y., Kong, T., Xu, C., Zhan, W., Tomizuka, M., Li, L., Yuan, Z., and Wang, C. (2021, January 19\u201325). Sparse R-CNN: End-to-end object detection with learnable proposals. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Nashville, TN, USA.","DOI":"10.1109\/CVPR46437.2021.01422"},{"key":"ref_56","first-page":"1870","article-title":"A fast and power-efficient hardware architecture for non-maximum suppression","volume":"66","author":"Shi","year":"2019","journal-title":"IEEE Trans. Circuits Syst. II Express Briefs"},{"key":"ref_57","doi-asserted-by":"crossref","unstructured":"Fang, C., Derbyshire, H., Sun, W., Yue, J., Shi, H., and Liu, Y. (2021, January 7\u201310). A sort-less FPGA-based non-maximum suppression accelerator using multi-thread computing and binary max engine for object detection. Proceedings of the IEEE Asian Solid-State Circuits Conference (A-SSCC), Busan, Republic of Korea.","DOI":"10.1109\/A-SSCC53895.2021.9634708"},{"key":"ref_58","doi-asserted-by":"crossref","unstructured":"Chen, C., Zhang, T., Yu, Z., Raghuraman, A., Udayan, S., Lin, J., and Aly, M.M.S. (2022, January 21\u201325). Scalable hardware acceleration of non-maximum suppression. Proceedings of the Design, Automation & Test in Europe Conference & Exhibition (DATE), Grenoble, France.","DOI":"10.23919\/DATE54114.2022.9774717"},{"key":"ref_59","doi-asserted-by":"crossref","unstructured":"Choi, S.B., Lee, S.S., Park, J., and Jang, S.J. (2021, January 1\u20133). Standard greedy non maximum suppression optimization for efficient and high speed inference. Proceedings of the IEEE International Conference on Consumer Electronics-Asia (ICCE-Asia), Gangwon, Republic of Korea.","DOI":"10.1109\/ICCE-Asia53811.2021.9641977"},{"key":"ref_60","unstructured":"Anupreetham, A., Ibrahim, M., Hall, M., Boutros, A., Kuzhively, A., Mohanty, A., Nurvitadhi, E., Betz, V., Cao, Y., and Seo, J. (September, January 30). End-to-end FPGA-based object detection using pipelined CNN and non-maximum suppression. Proceedings of the 31st International Conference on Field-Programmable Logic and Applications (FPL), Dresden, Germany."},{"key":"ref_61","doi-asserted-by":"crossref","unstructured":"Guo, Z., Liu, K., Liu, W., and Li, S. (2023, January 11\u201314). Efficient FPGA-based Accelerator for Post-Processing in Object Detection. Proceedings of the International Conference on Field Programmable Technology (ICFPT), Yokohama, Japan.","DOI":"10.1109\/ICFPT59805.2023.00019"},{"key":"ref_62","doi-asserted-by":"crossref","unstructured":"Chen, Y., Zhang, J., Lv, D., Yu, X., and He, G. (2023, January 21\u201325). O3 NMS: An Out-Of-Order-Based Low-Latency Accelerator for Non-Maximum Suppression. Proceedings of the IEEE International Symposium on Circuits and Systems (ISCAS), Monterey, CA, USA.","DOI":"10.1109\/ISCAS46773.2023.10181731"},{"key":"ref_63","first-page":"2251","article-title":"An area-efficient accelerator for non-maximum suppression","volume":"70","author":"Sun","year":"2023","journal-title":"IEEE Trans. Circuits Syst. II Express Briefs"},{"key":"ref_64","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/3634919","article-title":"High Throughput FPGA-Based Object Detection via Algorithm-Hardware Co-Design","volume":"17","author":"Anupreetham","year":"2024","journal-title":"ACM Trans. Reconfigurable Technol. Syst."},{"key":"ref_65","unstructured":"Batcher, K.E. (2, January April). Sorting networks and their applications. Proceedings of the Spring Joint Computer Conference, Atlantic City, NJ, USA."},{"key":"ref_66","doi-asserted-by":"crossref","first-page":"717","DOI":"10.1109\/TCSI.2023.3342929","article-title":"A Low-Cost Pipelined Architecture Based on a Hybrid Sorting Algorithm","volume":"71","author":"Chen","year":"2023","journal-title":"IEEE Trans. Circuits Syst. I Regul. Pap."},{"key":"ref_67","doi-asserted-by":"crossref","first-page":"2783","DOI":"10.1109\/TCAD.2024.3373592","article-title":"Hardware-software co-design enabling static and dynamic sparse attention mechanisms","volume":"43","author":"Zhao","year":"2024","journal-title":"IEEE Trans. Comput. Aided Des. Integr. Circuits Syst."},{"key":"ref_68","doi-asserted-by":"crossref","first-page":"506","DOI":"10.1109\/TCAD.2023.3317789","article-title":"Efficient N: M Sparse DNN Training Using Algorithm, Architecture, and Dataflow Co-Design","volume":"43","author":"Fang","year":"2023","journal-title":"IEEE Trans. Comput. Aided Des. Integr. Circuits Syst."},{"key":"ref_69","doi-asserted-by":"crossref","unstructured":"Zhang, H., Wu, W., Ma, Y., and Wang, Z. (2020, January 6\u20138). Efficient hardware post processing of anchor-based object detection on FPGA. Proceedings of the IEEE Computer Society Annual Symposium on VLSI (ISVLSI), Limassol, Cyprus.","DOI":"10.1109\/ISVLSI49217.2020.00089"}],"container-title":["Information"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/2078-2489\/16\/1\/63\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,8]],"date-time":"2025-10-08T10:31:01Z","timestamp":1759919461000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/2078-2489\/16\/1\/63"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,1,17]]},"references-count":69,"journal-issue":{"issue":"1","published-online":{"date-parts":[[2025,1]]}},"alternative-id":["info16010063"],"URL":"https:\/\/doi.org\/10.3390\/info16010063","relation":{},"ISSN":["2078-2489"],"issn-type":[{"value":"2078-2489","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,1,17]]}}}