{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,8,25]],"date-time":"2026-08-25T02:40:39Z","timestamp":1787625639767,"version":"build-2736575974"},"reference-count":57,"publisher":"Springer Science and Business Media LLC","issue":"5","license":[{"start":{"date-parts":[[2025,2,21]],"date-time":"2025-02-21T00:00:00Z","timestamp":1740096000000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/www.springernature.com\/gp\/researchers\/text-and-data-mining"},{"start":{"date-parts":[[2025,2,21]],"date-time":"2025-02-21T00:00:00Z","timestamp":1740096000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.springernature.com\/gp\/researchers\/text-and-data-mining"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Mach. Intell. Res."],"published-print":{"date-parts":[[2025,10]]},"DOI":"10.1007\/s11633-024-1508-2","type":"journal-article","created":{"date-parts":[[2025,2,20]],"date-time":"2025-02-20T22:00:51Z","timestamp":1740088851000},"page":"956-968","update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":5,"title":["LiDAR-camera Cooperative Semantic Segmentation"],"prefix":"10.1007","volume":"22","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-8827-5935","authenticated-orcid":false,"given":"He","family":"Guan","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-1223-3242","authenticated-orcid":false,"given":"Chunfeng","family":"Song","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-2648-3875","authenticated-orcid":false,"given":"Zhaoxiang","family":"Zhang","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2025,2,21]]},"reference":[{"issue":"7","key":"1508_CR1","doi-asserted-by":"publisher","first-page":"6063","DOI":"10.1109\/TITS.2021.3076844","volume":"23","author":"B Gao","year":"2022","unstructured":"B. Gao, Y. C Pan, C. K. Li, S. B. Geng, H. J. Zhao. Are we hungry for 3D LiDAR data for semantic segmentation? A survey of datasets and methods. IEEE Transactions on Intelligent Transportation Systems, vol. 23, no. 7, pp. 6063\u20136081, 2022. DOI: https:\/\/doi.org\/10.1109\/TITS.2021.3076844.","journal-title":"IEEE Transactions on Intelligent Transportation Systems"},{"key":"1508_CR2","doi-asserted-by":"publisher","unstructured":"G. Rizzoli, F. Barbato, P. Zanuttigh. Multimodal semantic segmentation in autonomous driving: A review of current approaches and future perspectives. Technologies, vol. 10, no. 4, Article number 90, 2022. DOI: https:\/\/doi.org\/10.3390\/technologies10040090.","DOI":"10.3390\/technologies10040090"},{"key":"1508_CR3","volume-title":"Multi-modal sensor fusion for auto driving perception: A survey","author":"K L Huang","year":"2024","unstructured":"K. L. Huang, B. T. Shi, X. Li, X. Li, S. Y. Huang, Y. K. Li. Multi-modal sensor fusion for auto driving perception: A survey, [Online], Available: http:\/\/arxiv.org\/abs\/2202.02703, 2024."},{"key":"1508_CR4","doi-asserted-by":"publisher","first-page":"1743","DOI":"10.1109\/CVPR.2017.189","volume-title":"Proceedings of IEEE Conference on Computer Vision and Pattern Recognition","author":"C Peng","year":"2017","unstructured":"C. Peng, X. Y. Zhang, G. Yu, G. M. Luo, J. Sun. Large kernel matters-improve semantic segmentation by global convolutional network. In Proceedings of IEEE Conference on Computer Vision and Pattern Recognition, Honolulu, USA, pp. 1743\u20131751, 2017. DOI: https:\/\/doi.org\/10.1109\/CVPR.2017.189."},{"key":"1508_CR5","doi-asserted-by":"publisher","first-page":"6230","DOI":"10.1109\/CVPR.2017.660","volume-title":"Proceedings of IEEE Conference on Computer Vision and Pattern Recognition","author":"H S Zhao","year":"2017","unstructured":"H. S. Zhao, J. P. Shi, X. J. Qi, X. G. Wang, J. Y. Jia. Pyramid scene parsing network. In Proceedings of IEEE Conference on Computer Vision and Pattern Recognition, Honolulu, USA, pp. 6230\u20136239, 2017. DOI: https:\/\/doi.org\/10.1109\/CVPR.2017.660."},{"key":"1508_CR6","doi-asserted-by":"publisher","first-page":"603","DOI":"10.1109\/ICCV.2019.00069","volume-title":"Proceedings of IEEE\/CVF International Conference on Computer Vision","author":"Z L Huang","year":"2019","unstructured":"Z. L. Huang, X. G. Wang, L. C Huang, C Huang, Y. C Wei, W. Y. Liu. CCNet: Criss-cross attention for semantic segmentation. In Proceedings of IEEE\/CVF International Conference on Computer Vision, Seoul, Republic of Korea, pp. 603\u2013612, 2019. DOI: https:\/\/doi.org\/10.1109\/ICCV.2019.00069."},{"key":"1508_CR7","doi-asserted-by":"publisher","first-page":"82","DOI":"10.1109\/CVPR.2019.00017","volume-title":"Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"C X Liu","year":"2019","unstructured":"C. X. Liu, L. C. Chen, F. Schroff, H. Adam, W. Hua, A. L. Yuille, L. Fei-Fei. Auto-DeepLab: Hierarchical neural architecture search for semantic image segmentation. In Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Long Beach, USA, pp. 82\u201392, 2019. DOI: https:\/\/doi.org\/10.1109\/CVPR.2019.00017."},{"key":"1508_CR8","first-page":"12077","volume-title":"Proceedings of the 35th Conference on Neural Information Processing Systems","author":"E Z Xie","year":"2021","unstructured":"E. Z. Xie, W. H. Wang, Z. D. Yu, A. Anandkumar, J. M. \u00c1lvarez, P. Luo. SegFormer: Simple and efficient design for semantic segmentation with transformers. In Proceedings of the 35th Conference on Neural Information Processing Systems, Montreal, Canada, pp. 12077\u201312090, 2021."},{"key":"1508_CR9","doi-asserted-by":"publisher","first-page":"12073","DOI":"10.1109\/CVPR52688.2022.01177","volume-title":"Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"W Q Zhang","year":"2022","unstructured":"W. Q. Zhang, Z. L. Huang, G. Z. Luo, T. Chen, X. G. Wang, W. Y. Liu, G. Yu, C. H. Shen TopFormer: Token pyramid transformer for mobile semantic segmentation In Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition, New Orleans, USA, pp 12073\u201312083, 2022. DOI: https:\/\/doi.org\/10.1109\/CVPR52688.2022.01177."},{"key":"1508_CR10","doi-asserted-by":"publisher","first-page":"11105","DOI":"10.1109\/CVPR42600.2020.01112","volume-title":"Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"Q Y Hu","year":"2020","unstructured":"Q. Y Hu, B. Yang, L. H. Xie, S. Rosa, Y. L. Guo, Z. H. Wang, N. Trigoni, A. Markham. RandLA-Net: Efficient semantic segmentation of large-scale point clouds. In Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Seattle, USA, pp. 11105\u201311114, 2020. DOI: https:\/\/doi.org\/10.1109\/CVPR42600.2020.01112."},{"key":"1508_CR11","doi-asserted-by":"publisher","first-page":"685","DOI":"10.1007\/978-3-030-58604-1_41","volume-title":"Proceedings of the 16th European Conference on Computer Vision","author":"H T Tang","year":"2020","unstructured":"H. T. Tang, Z. J. Liu, S. Y. Zhao, Y. J. Lin, J. Lin, H. R. Wang, S. Han. Searching efficient 3D architectures with sparse point-voxel convolution. In Proceedings of the 16th European Conference on Computer Vision, Glasgow, UK, pp. 685\u2013702, 2020. DOI: https:\/\/doi.org\/10.1007\/978-3-030-58604-1_41."},{"key":"1508_CR12","first-page":"9939","volume-title":"Proceedings of IEEE Conference on Computer Vision and Pattern Recognition","author":"X G Zhu","year":"2021","unstructured":"X. G. Zhu, H. Zhou, T. Wang, F. Z. Hong, Y. X. Ma, W. Li, H. S. Li, D. H. Lin. Cylindrical and asymmetrical 3D convolution networks for LiDAR segmentation. In Proceedings of IEEE Conference on Computer Vision and Pattern Recognition, Nashville, USA, pp. 9939\u20139948, 2021."},{"key":"1508_CR13","doi-asserted-by":"publisher","first-page":"207","DOI":"10.1007\/978-3-030-64559-5_16","volume-title":"Proceedings of the 15th International Symposium Advances in Visual Computing","author":"T Cortinhal","year":"2020","unstructured":"T. Cortinhal, G. Tzelepis, E. Erdal Aksoy. SalsaNext: Fast, uncertainty-aware semantic segmentation of LiDAR point clouds. In Proceedings of the 15th International Symposium Advances in Visual Computing, San Diego, USA, pp.207\u2013222, 2020. DOI: https:\/\/doi.org\/10.1007\/978-3-030-64559-5_16."},{"issue":"4","key":"1508_CR14","doi-asserted-by":"publisher","first-page":"834","DOI":"10.1109\/TPAMI.2017.2699184","volume":"40","author":"L C Chen","year":"2018","unstructured":"L. C. Chen, G. Papandreou, I. Kokkinos, K. Murphy, A. L. Yuille. DeepLab: Semantic image segmentation with deep convolutional nets, atrous convolution, and fully connected CRFs. IEEE Transactions on Pattern Analysis and Machine Intelligence, vol. 40, no. 4, pp. 834\u2013848, 2018. DOI: https:\/\/doi.org\/10.1109\/TPAMI.2017.2699184.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"1508_CR15","doi-asserted-by":"publisher","first-page":"801","DOI":"10.1007\/978-3-030-01234-2_49","volume-title":"Proceedings of the 15th European Conference on Computer Vision","author":"L C Chen","year":"2018","unstructured":"L. C. Chen, Y. K. Zhu, G. Papandreou, F. Schroff, H. Adam. Encoder-decoder with atrous separable convolution for semantic image segmentation. In Proceedings of the 15th European Conference on Computer Vision, Munich, Germany, pp. 801\u2013818, 2018. DOI: https:\/\/doi.org\/10.1007\/978-3-030-01234-2_49."},{"key":"1508_CR16","doi-asserted-by":"publisher","first-page":"12413","DOI":"10.1109\/CVPR42600.2020.01243","volume-title":"Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"C Q Yu","year":"2020","unstructured":"C. Q. Yu, J. B. Wang, C. X. Gao, G. Yu, C. H. Shen, N. Sang. Context prior for scene segmentation. In Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Seattle, USA, pp. 12413\u201312422, 2020. DOI: https:\/\/doi.org\/10.1109\/CVPR42600.2020.01243."},{"key":"1508_CR17","doi-asserted-by":"publisher","first-page":"7151","DOI":"10.1109\/CVPR.2018.00747","volume-title":"Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"H Zhang","year":"2018","unstructured":"H. Zhang, K. Dana, J. P. Shi, Z. Y. Zhang, X. G. Wang, A. Tyagi, A. Agrawal. Context encoding for semantic segmentation. In Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Salt Lake City, USA, pp. 7151\u20137160, 2018. DOI: https:\/\/doi.org\/10.1109\/CVPR.2018.00747."},{"key":"1508_CR18","doi-asserted-by":"publisher","first-page":"13663","DOI":"10.1109\/CVPR42600.2020.01368","volume-title":"Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"M M Zhen","year":"2020","unstructured":"M. M. Zhen, J. L. Wang, L. Zhou, S. W. Li, T. W. Shen, J. X. Shang, T. Fang, L. Quan. Joint semantic segmentation and boundary detection using iterative pyramid contexts. In Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Seattle, USA, pp. 13663\u201313672, 2020. DOI: https:\/\/doi.org\/10.1109\/CVPR42600.2020.01368."},{"key":"1508_CR19","doi-asserted-by":"publisher","first-page":"6818","DOI":"10.1109\/ICCV.2019.00692","volume-title":"Proceedings of IEEE\/CVF International Conference on Computer Vision","author":"H H Ding","year":"2019","unstructured":"H. H. Ding, X. D. Jiang, A. Q. Liu, N. M. Thalmann, G. Wang Boundary-aware feature propagation for scene segmentation In Proceedings of IEEE\/CVF International Conference on Computer Vision, Seoul, Republic of Korea, pp. 6818\u20136828, 2019. DOI: https:\/\/doi.org\/10.1109\/ICCV.2019.00692."},{"key":"1508_CR20","doi-asserted-by":"publisher","first-page":"2014","DOI":"10.1109\/ICCVW.2019.00251","volume-title":"Proceedings of IEEE\/CVF International Conference on Computer Vision Workshops","author":"A Shaw","year":"2019","unstructured":"A. Shaw, D. Hunter, F. Landola, S. Sidhu. SqueezeNAS: Fast neural architecture search for faster semantic segmentation. In Proceedings of IEEE\/CVF International Conference on Computer Vision Workshops, Seoul, Republic of Korea, pp. 2014\u20132024, 2019. DOI: https:\/\/doi.org\/10.1109\/ICCVW.2019.00251."},{"key":"1508_CR21","doi-asserted-by":"publisher","first-page":"2989","DOI":"10.1109\/CVPR52729.2023.00292","volume-title":"Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"J Jain","year":"2023","unstructured":"J. Jain, J. C. Li, M. T. Chiu, A. Hassani, N. Orlov, H. Shi. OneFormer: One transformer to rule universal image segmentation. In Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Vancouver, Canada, pp. 2989\u20132998, 2023. DOI: https:\/\/doi.org\/10.1109\/CVPR52729.2023.00292."},{"key":"1508_CR22","doi-asserted-by":"publisher","first-page":"77","DOI":"10.1109\/CVPR.2017.16","volume-title":"Proceedings of IEEE Conference on Computer Vision and Pattern Recognition","author":"R Q Charles","year":"2017","unstructured":"R. Q. Charles, H. Su, M. Kaichun, L. J. Guibas. PointNet: Deep learning on point sets for 3D classification and segmentation. In Proceedings of IEEE Conference on Computer Vision and Pattern Recognition, Honolulu, USA, pp. 77\u201385, 2017. DOI: https:\/\/doi.org\/10.1109\/CVPR.2017.16."},{"key":"1508_CR23","first-page":"5105","volume-title":"Proceedings of the 31st International Conference on Neural Information Processing Systems","author":"R Q Charles","year":"2017","unstructured":"R. Q. Charles, L. Yi, H. Su, L. J. Guibas. PointNet++: Deep hierarchical feature learning on point sets in a metric space. In Proceedings of the 31st International Conference on Neural Information Processing Systems, Long Beach, USA, pp. 5105\u20135114, 2017."},{"key":"1508_CR24","doi-asserted-by":"publisher","unstructured":"Y. Wang, Y. B. Sun, Z. W. Liu, S. E. Sarma, M. M. Bronstein, J. M. Solomon. Dynamic graph CNN for learning on point clouds. ACM Transactions on Graphics, vol. 38, no. 5, Article number 146, 2019. DOI: https:\/\/doi.org\/10.1145\/3326362.","DOI":"10.1145\/3326362"},{"key":"1508_CR25","doi-asserted-by":"publisher","first-page":"9613","DOI":"10.1109\/CVPR.2019.00985","volume-title":"Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"W X Wu","year":"2019","unstructured":"W. X. Wu, Z. G. Qi, F. X. Li. PointConv: Deep convolutional networks on 3D point clouds. In Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Long Beach, USA, pp. 9613\u20139622, 2019. DOI: https:\/\/doi.org\/10.1109\/CVPR.2019.00985."},{"key":"1508_CR26","doi-asserted-by":"publisher","first-page":"6410","DOI":"10.1109\/ICCV.2019.00651","volume-title":"Proceedings of IEEE\/CVF International Conference on Computer Vision","author":"H Thomas","year":"2019","unstructured":"H. Thomas, C. R. Qi, J. E. Deschaud, B. Marcotegui, F. Goulette, L. J. Guibas. KPConv: Flexible and deformable convolution for point clouds. In Proceedings of IEEE\/CVF International Conference on Computer Vision, Seoul, Republic of Korea, pp. 6410\u20136419, 2019. DOI: https:\/\/doi.org\/10.1109\/ICCV.2019.00651."},{"key":"1508_CR27","doi-asserted-by":"publisher","first-page":"984","DOI":"10.1109\/CVPR.2018.00109","volume-title":"Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"B S Hua","year":"2018","unstructured":"B. S. Hua, M. K. Tran, S. K. Yeung. Pointwise convolutional neural networks. In Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Salt Lake City, USA, pp. 984\u2013993, 2018. DOI: https:\/\/doi.org\/10.1109\/CVPR.2018.00109."},{"key":"1508_CR28","doi-asserted-by":"publisher","first-page":"95","DOI":"10.1007\/978-3-319-64689-3_8","volume-title":"Proceedings of the 17th International Conference on Computer Analysis of Images and Patterns","author":"F J Lawin","year":"2017","unstructured":"F. J. Lawin, M. Danelljan, P. Tosteberg, G. Bhat, F. S. Khan, M. Felsberg. Deep projective 3D semantic segmentation. In Proceedings of the 17th International Conference on Computer Analysis of Images and Patterns, Ystad, Sweden, pp. 95\u2013107, 2017. DOI: https:\/\/doi.org\/10.1007\/978-3-319-64689-3_8."},{"key":"1508_CR29","doi-asserted-by":"publisher","first-page":"3887","DOI":"10.1109\/CVPR.2018.00409","volume-title":"Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"M Tatarchenko","year":"2018","unstructured":"M. Tatarchenko, J. Park, V. Koltun, Q. Y. Zhou. Tangent convolutions for dense prediction in 3D. In Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Salt Lake City, USA, pp. 3887\u20133896, 2018. DOI: https:\/\/doi.org\/10.1109\/CVPR.2018.00409."},{"key":"1508_CR30","doi-asserted-by":"publisher","first-page":"1887","DOI":"10.1109\/ICRA.2018.8462926","volume-title":"Proceedings of IEEE International Conference on Robotics and Automation","author":"B C Wu","year":"2018","unstructured":"B. C. Wu, A. Wan, X. Y. Yue, K. Keutzer. SqueezeSeg: Convolutional neural nets with recurrent CRF for realtime road-object segmentation from 3D LiDAR point cloud. In Proceedings of IEEE International Conference on Robotics and Automation, Brisbane, Australia, pp. 1887\u20131893, 2018. DOI: https:\/\/doi.org\/10.1109\/ICRA.2018.8462926."},{"key":"1508_CR31","doi-asserted-by":"publisher","first-page":"4376","DOI":"10.1109\/ICRA.2019.8793495","volume-title":"Proceedings of International Conference on Robotics and Automation","author":"B C Wu","year":"2019","unstructured":"B. C. Wu, X. Y. Zhou, S. C. Zhao, X. Y. Yue, K. Keutzer. SqueezeSegV2: Improved model structure and unsupervised domain adaptation for road-object segmentation from a LiDAR point cloud. In Proceedings of International Conference on Robotics and Automation, Montreal, Canada, pp. 4376\u20134382, 2019. DOI: https:\/\/doi.org\/10.1109\/ICRA.2019.8793495."},{"key":"1508_CR32","volume-title":"AMVNet: Assertion-based multi-view fusion network for LiDAR semantic segmentation","author":"V E Liong","year":"2023","unstructured":"V. E. Liong, T. H. T. Nguyen, S. Widjaja, D. Sharma, Z. J. Chong. AMVNet: Assertion-based multi-view fusion network for LiDAR semantic segmentation, [Online], Available: https:\/\/arxiv.org\/abs\/2012.04934, 2023."},{"key":"1508_CR33","doi-asserted-by":"publisher","first-page":"9224","DOI":"10.1109\/CVPR.2018.00961","volume-title":"Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"B Graham","year":"2018","unstructured":"B. Graham, M. Engelcke, L. Van Der Maaten. 3D semantic segmentation with submanifold sparse convolutional networks. In Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Salt Lake City, USA, pp. 9224\u20139232, 2018. DOI: https:\/\/doi.org\/10.1109\/CVPR.2018.00961."},{"key":"1508_CR34","doi-asserted-by":"publisher","first-page":"9598","DOI":"10.1109\/CVPR42600.2020.00962","volume-title":"Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"Y Zhang","year":"2020","unstructured":"Y. Zhang, Z. X. Zhou, P. David, X. Y. Yue, Z. R. Xi, B. Q. Gong, H. Foroosh. PolarNet: An improved grid representation for online LiDAR point clouds semantic segmentation. In Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Seattle, USA, pp. 9598\u20139607, 2020. DOI: https:\/\/doi.org\/10.1109\/CVPR42600.2020.00962."},{"key":"1508_CR35","doi-asserted-by":"publisher","first-page":"17545","DOI":"10.1109\/CVPR52729.2023.01683","volume-title":"Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"X Lai","year":"2023","unstructured":"X. Lai, Y. K. Chen, F. B. Lu, J. H. Liu, J. Y. Jia. Spherical transformer for LiDAR-based 3D recognition. In Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Vancouver, Canada, pp. 17545\u201317555, 2023. DOI: https:\/\/doi.org\/10.1109\/CVPR52729.2023.01683."},{"key":"1508_CR36","doi-asserted-by":"publisher","first-page":"14040","DOI":"10.1109\/ICRA48506.2021.9561305","volume-title":"Proceedings of IEEE International Conference on Robotics and Automation","author":"R Cheng","year":"2021","unstructured":"R. Cheng, R. Razani, Y. Ren, B. B. Liu. S3Net: 3D LiDAR sparse semantic segmentation network. In Proceedings of IEEE International Conference on Robotics and Automation, Xi\u2019an, China, pp. 14040\u201314046, 2021. DOI: https:\/\/doi.org\/10.1109\/ICRA48506.2021.9561305."},{"key":"1508_CR37","doi-asserted-by":"publisher","first-page":"16004","DOI":"10.1109\/ICCV48922.2021.01572","volume-title":"Proceedings of IEEE\/CVF International Conference on Computer Vision","author":"J Y Xu","year":"2021","unstructured":"J. Y. Xu, R. X. Zhang, J. Dou, Y. S. Zhu, J. Sun, S. L. Pu. RPVNet: A deep and efficient range-point-voxel fusion network for LiDAR point cloud segmentation. In Proceedings of IEEE\/CVF International Conference on Computer Vision, Montreal, Canada, pp. 16004\u201316013, 2021. DOI: https:\/\/doi.org\/10.1109\/ICCV48922.2021.01572."},{"key":"1508_CR38","doi-asserted-by":"publisher","first-page":"16260","DOI":"10.1109\/ICCV48922.2021.01597","volume-title":"Proceedings of IEEE\/CVF International Conference on Computer Vision","author":"Z W Zhuang","year":"2021","unstructured":"Z. W. Zhuang, R. Li, K. Jia, Q. C. Wang, Y. Q. Li, M. K. Tan. Perception-aware multi-sensor fusion for 3D LiDAR semantic segmentation. In Proceedings of IEEE\/CVF International Conference on Computer Vision, Montreal, Canada, pp. 16260\u201316270, 2021. DOI: https:\/\/doi.org\/10.1109\/ICCV48922.2021.01597."},{"key":"1508_CR39","doi-asserted-by":"publisher","first-page":"7","DOI":"10.1109\/ITSC.2019.8917447","volume-title":"IEEE Intelligent Transportation Systems Conference","author":"K El Madawi","year":"2019","unstructured":"K. El Madawi, H. Rashed, A. El Sallab, O. Nasr, H. Kamel, S. Yogamani. RGB and LiDAR fusion based 3D semantic segmentation for autonomous driving. In IEEE Intelligent Transportation Systems Conference, Auckland, New Zealand, pp. 7\u201312, 2019. DOI: https:\/\/doi.org\/10.1109\/ITSC.2019.8917447."},{"key":"1508_CR40","doi-asserted-by":"publisher","first-page":"1863","DOI":"10.1109\/WACV45572.2020.9093584","volume-title":"Proceedings of IEEE Winter Conference on Applications of Computer Vision","author":"G Krispel","year":"2020","unstructured":"G. Krispel, M. Opitz, G. Waltner, H. Possegger, H. Bischof. FuseSeg: LiDAR point cloud segmentation fusing multi-modal data. In Proceedings of IEEE Winter Conference on Applications of Computer Vision, Snowmass, USA, pp. 1863\u20131872, 2020. DOI: https:\/\/doi.org\/10.1109\/WACV45572.2020.9093584."},{"key":"1508_CR41","doi-asserted-by":"publisher","unstructured":"S. Vora, A. H. Lang, B. Helou, O. Beijbom. PointPainting: Sequential fusion for 3d object detection. In Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Seattle, USA, pp. 4603\u20134611. DOI: https:\/\/doi.org\/10.1109\/CVPR42600.2020.00466.","DOI":"10.1109\/CVPR42600.2020.00466"},{"key":"1508_CR42","doi-asserted-by":"publisher","first-page":"677","DOI":"10.1007\/978-3-031-19815-1_39","volume-title":"Proceedings of the 17th European Conference on Computer Vision","author":"X Yan","year":"2022","unstructured":"X. Yan, J. T. Gao, C. D. Zheng, C. Zheng, R. M. Zhang, S. G. Cui, Z. Li. 2DPASS: 2D priors assisted semantic segmentation on LiDAR point clouds. In Proceedings of the 17th European Conference on Computer Vision, Tel Aviv-Yafo, Israel, pp. 677\u2013695, 2022. DOI: https:\/\/doi.org\/10.1007\/978-3-031-19815-1_39."},{"key":"1508_CR43","doi-asserted-by":"publisher","first-page":"11789","DOI":"10.1109\/CVPR46437.2021.01162","volume-title":"Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"C W Wang","year":"2021","unstructured":"C. W. Wang, C. Ma, M. Zhu, X. K. Yang. PointAugmenting: Cross-modal augmentation for 3D object detection. In Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Nashville, USA, pp. 11789\u201311798, 2021. DOI: https:\/\/doi.org\/10.1109\/CVPR46437.2021.01162."},{"key":"1508_CR44","doi-asserted-by":"publisher","first-page":"5408","DOI":"10.1109\/CVPR52688.2022.00534","volume-title":"Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"X P Wu","year":"2022","unstructured":"X. P. Wu, L. Peng, H. H. Yang, L. Xie, C. X. Huang, C. Q. Deng, H. F. Liu, D. Cai. Sparse fuse dense: Towards high quality 3D detection with depth completion. In Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition, New Orleans, USA, USA, pp. 5408\u20135417, 2022. DOI: https:\/\/doi.org\/10.1109\/CVPR52688.2022.00534."},{"key":"1508_CR45","doi-asserted-by":"publisher","first-page":"1080","DOI":"10.1109\/CVPR52688.2022.00116","volume-title":"Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"X Y Bai","year":"2022","unstructured":"X. Y. Bai, Z. Y. Hu, X. G. Zhu, Q. Q. Huang, Y. L. Chen, H. B. Fu, C. L. Tai. TransFusion: Robust LiDAR-camera fusion for 3D object detection with transformers. In Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition, New Orleans, USA, pp. 1080\u20131089, 2022. DOI: https:\/\/doi.org\/10.1109\/CVPR52688.2022.00116."},{"key":"1508_CR46","doi-asserted-by":"publisher","first-page":"11966","DOI":"10.1109\/CVPR52688.2022.01167","volume-title":"Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"Z Liu","year":"2022","unstructured":"Z. Liu, H. Z. Mao, C. Y. Wu, C. Feichtenhofer, T. Darrell, S. N. Xie. A convNet for the 2020s. In Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition, New Orleans, USA, pp. 11966\u201311976, 2022. DOI: https:\/\/doi.org\/10.1109\/CVPR52688.2022.01167."},{"key":"1508_CR47","first-page":"180","volume-title":"Proceedings of the 5th Conference on Robot Learning","author":"Y Wang","year":"2021","unstructured":"Y. Wang, V. Guizilini, T. Y. Zhang, Y. L. Wang, H. Zhao, J. Solomon. DETR3D: 3D object detection from multiview images via 3D-to-2D queries. In Proceedings of the 5th Conference on Robot Learning, London, UK, pp. 180\u2013191, 2021."},{"key":"1508_CR48","first-page":"302","volume-title":"Proceedings of the 5th Machine Learning and Systems","author":"H T Tang","year":"2022","unstructured":"H. T. Tang, Z. J. Liu, X. Y. Li, Y. J. Lin, S. Han. Torchsparse: Efficient point cloud inference engine. In Proceedings of the 5th Machine Learning and Systems, Santa Clara, USA, pp. 302\u2013315, 2022."},{"key":"1508_CR49","doi-asserted-by":"publisher","first-page":"2443","DOI":"10.1109\/CVPR42600.2020.00252","volume-title":"Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"P Sun","year":"2020","unstructured":"P. Sun, H. Kretzschmar, X. Dotiwalla, A. Chouard, V. Patnaik, P. Tsui, J. Guo, Y. Zhou, Y. N. Chai, B. Caine, V. Vasudevan, W. Han, J. Ngiam, H. Zhao, A. Timofeev, S. Ettinger, M. Krivokon, A. Gao, A. Joshi, Y. Zhang, J. Shlens, Z. F. Chen, D. Anguelov. Scalability in perception for autonomous driving: Waymo open dataset. In Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Seattle, USA, pp. 2443\u20132451, 2020. DOI: https:\/\/doi.org\/10.1109\/CVPR42600.2020.00252."},{"key":"1508_CR50","doi-asserted-by":"publisher","first-page":"9296","DOI":"10.1109\/ICCV.2019.00939","volume-title":"Proceedings of IEEE\/CVF International Conference on Computer Vision","author":"J Behley","year":"2019","unstructured":"J. Behley, M. Garbade, A. Milioto, J. Quenzel, S. Behnke, C. Stachniss, J. Gall. SemanticKITTI: A dataset for semantic scene understanding of LiDAR sequences. In Proceedings of IEEE\/CVF International Conference on Computer Vision, Seoul, Republic of Korea, pp.9296\u20139306, 2019. DOI: https:\/\/doi.org\/10.1109\/ICCV.2019.00939."},{"key":"1508_CR51","doi-asserted-by":"publisher","first-page":"3070","DOI":"10.1109\/CVPR.2019.00319","volume-title":"Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"C Choy","year":"2019","unstructured":"C. Choy, J. Y. Gwak, S. Savarese. 4D spatio-temporal convNets: Minkowski convolutional neural networks. In Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Long Beach, USA, pp. 3070\u20133079, 2019. DOI: https:\/\/doi.org\/10.1109\/CVPR.2019.00319."},{"key":"1508_CR52","doi-asserted-by":"publisher","first-page":"770","DOI":"10.1109\/CVPR.2016.90","volume-title":"Proceedings of IEEE Conference on Computer Vision and Pattern Recognition","author":"K M He","year":"2016","unstructured":"K. M. He, X. Y. Zhang, S. Q. Ren, J. Sun. Deep residual learning for image recognition. In Proceedings of IEEE Conference on Computer Vision and Pattern Recognition, Las Vegas, USA, pp. 770\u2013778, 2016. DOI: https:\/\/doi.org\/10.1109\/CVPR.2016.90."},{"key":"1508_CR53","doi-asserted-by":"publisher","first-page":"21706","DOI":"10.1109\/CVPR52729.2023.02079","volume-title":"Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"L D Kong","year":"2023","unstructured":"L. D. Kong, J. W. Ren, L. Pan, Z. W. Liu. LaserMix for semi-supervised LiDAR semantic segmentation. In Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Vancouver, Canada, pp. 21706\u201321716, 2023. DOI: https:\/\/doi.org\/10.1109\/CVPR52729.2023.02079."},{"key":"1508_CR54","first-page":"11035","volume-title":"Proceedings of the 36th Conference on Neural Information Processing Systems","author":"A R Xiao","year":"2022","unstructured":"A. R. Xiao, J. X. Huang, D. Y. Guan, K. W. Cui, S. J. Lu, L. Shao. PolarMix: A general data augmentation technique for LiDAR point clouds. In Proceedings of the 36th Conference on Neural Information Processing Systems, New Orleans, USA, pp. 11035\u201311048, 2022."},{"key":"1508_CR55","volume-title":"Cylinder3D: An effective 3D framework for driving-scene LiDAR semantic segmentation","author":"H Zhou","year":"2023","unstructured":"H. Zhou, X. G. Zhu, X. Song, Y. X. Ma, Z. Wang, H. S. Li, D. H. Lin. Cylinder3D: An effective 3D framework for driving-scene LiDAR semantic segmentation, [Online], Available: https:\/\/arxiv.org\/abs\/2008.01550, 2023."},{"key":"1508_CR56","doi-asserted-by":"publisher","first-page":"11779","DOI":"10.1109\/CVPR46437.2021.01161","volume-title":"Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"T W Yin","year":"2021","unstructured":"T. W. Yin, X. Y. Zhou, P. Kr\u00e4henb\u00fchl. Center-based 3D object detection and tracking. In Proceedings of IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Nashville, USA, pp. 11779\u201311788, 2021. DOI: https:\/\/doi.org\/10.1109\/CVPR46437.2021.01161."},{"key":"1508_CR57","first-page":"10096","volume-title":"Proceedings of the 38th International Conference on Machine Learning","author":"M X Tan","year":"2021","unstructured":"M. X. Tan, Q. V. Le. EfficientNetV2: Smaller models and faster training. In Proceedings of the 38th International Conference on Machine Learning, Vienna, Austria, pp. 10096\u201310106, 2021."}],"container-title":["Machine Intelligence Research"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s11633-024-1508-2.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s11633-024-1508-2\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s11633-024-1508-2.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,9,27]],"date-time":"2025-09-27T08:02:41Z","timestamp":1758960161000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s11633-024-1508-2"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,2,21]]},"references-count":57,"journal-issue":{"issue":"5","published-print":{"date-parts":[[2025,10]]}},"alternative-id":["1508"],"URL":"https:\/\/doi.org\/10.1007\/s11633-024-1508-2","relation":{},"ISSN":["2731-538X","2731-5398"],"issn-type":[{"value":"2731-538X","type":"print"},{"value":"2731-5398","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,2,21]]},"assertion":[{"value":"17 July 2023","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"3 April 2024","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"21 February 2025","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"The authors declared that they have no conflicts of interest to this work.","order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations of conflict of interest"}}]}}