{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,8,14]],"date-time":"2026-08-14T15:53:46Z","timestamp":1786722826869,"version":"3.56.0"},"reference-count":59,"publisher":"Springer Science and Business Media LLC","issue":"2","license":[{"start":{"date-parts":[[2026,2,6]],"date-time":"2026-02-06T00:00:00Z","timestamp":1770336000000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2026,2,6]],"date-time":"2026-02-06T00:00:00Z","timestamp":1770336000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"name":"Centre for Research & Technology Hellas"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Machine Vision and Applications"],"published-print":{"date-parts":[[2026,3]]},"abstract":"<jats:title>Abstract<\/jats:title>\n                  <jats:p>AR applications are rapidly gaining adoption, expanding users\u2019 perceptual capabilities. Their potential for addressing visual occlusions and effectively extending the user\u2019s line of sight significantly enhances their value in various contexts, from professional to personal use. This work presents a novel system designed to project 3D human poses onto AR glasses, enabling users to perceive concealed individuals behind solid objects, addressing a critical limitation of traditional visual perception. To achieve real-time and accurate 3D projection, we employ fiducial markers strategically placed within the environment. The markers are periodically fused with IMU sensor data to accurately estimate the user\u2019s head orientation, a crucial step for correct spatial alignment. Furthermore, we leverage a multi-view 3D human pose estimation method using calibrated cameras and incorporate attention mechanisms. These mechanisms focus the system on relevant features, improving accuracy and minimizing 3D joint error. Our experiments demonstrate that the proposed framework accurately projects 3D skeletal representations onto AR glasses, even when significant occlusions are caused by solid objects or other occupants within the scene. This novel approach offers a method to enhance situational awareness in dynamic environments where visibility is compromised, potentially benefiting various applications, from first response scenarios to security and surveillance.<\/jats:p>","DOI":"10.1007\/s00138-025-01783-9","type":"journal-article","created":{"date-parts":[[2026,2,6]],"date-time":"2026-02-06T11:47:07Z","timestamp":1770378427000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":1,"title":["Overcoming occlusions in AR, via multi-view, real-time 3D human pose estimation"],"prefix":"10.1007","volume":"37","author":[{"ORCID":"https:\/\/orcid.org\/0009-0004-2859-467X","authenticated-orcid":false,"given":"Ioannis","family":"Pastaltzidis","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-4786-3060","authenticated-orcid":false,"given":"Iason","family":"Karakostas","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-6650-7758","authenticated-orcid":false,"given":"Nikolaos","family":"Dimitriou","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-9666-7023","authenticated-orcid":false,"given":"Stelios","family":"Krinidis","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-6915-6722","authenticated-orcid":false,"given":"Dimitrios","family":"Tzovaras","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2026,2,6]]},"reference":[{"issue":"10","key":"1783_CR1","doi-asserted-by":"publisher","first-page":"4031","DOI":"10.1109\/TVCG.2022.3175626","volume":"29","author":"CD Linhares","year":"2022","unstructured":"Linhares, C.D., Lima, D.M., Ponciano, J.R., Olivatto, M.M., Gutierrez, M.A., Poco, J., Traina, C., Traina, A.J.M.: Clinicalpath: a visualization tool to improve the evaluation of electronic health records in clinical decision-making. IEEE Trans. Visual Comput. Graphics 29(10), 4031\u20134046 (2022)","journal-title":"IEEE Trans. Visual Comput. Graphics"},{"issue":"12","key":"1783_CR2","doi-asserted-by":"publisher","first-page":"5062","DOI":"10.1109\/TVCG.2022.3201120","volume":"29","author":"IM Butaslac","year":"2022","unstructured":"Butaslac, I.M., Fujimoto, Y., Sawabe, T., Kanbara, M., Kato, H.: Systematic review of augmented reality training systems. IEEE Trans. Visual Comput. Graphics 29(12), 5062\u20135082 (2022)","journal-title":"IEEE Trans. Visual Comput. Graphics"},{"issue":"12","key":"1783_CR3","doi-asserted-by":"publisher","first-page":"4832","DOI":"10.1109\/TVCG.2022.3193672","volume":"29","author":"S Jadhav","year":"2022","unstructured":"Jadhav, S., Kaufman, A.E.: Md-cave: An immersive visualization workbench for radiologists. IEEE Trans. Visual Comput. Graphics 29(12), 4832\u20134844 (2022)","journal-title":"IEEE Trans. Visual Comput. Graphics"},{"issue":"11","key":"1783_CR4","doi-asserted-by":"publisher","first-page":"4655","DOI":"10.1109\/TVCG.2023.3320223","volume":"29","author":"J Sermarini","year":"2023","unstructured":"Sermarini, J., Michlowitz, R.A., LaViola, J.J., Walters, L.C., Azevedo, R., Kider, J.T.: Investigating the impact of augmented reality and bim on retrofitting training for non-experts. IEEE Trans. Visual Comput. Graphics 29(11), 4655\u20134665 (2023). https:\/\/doi.org\/10.1109\/TVCG.2023.3320223","journal-title":"IEEE Trans. Visual Comput. Graphics"},{"issue":"1","key":"1783_CR5","first-page":"62","volume":"6","author":"G Mathioudakis","year":"2014","unstructured":"Mathioudakis, G., Leonidis, A., Korozi, M., Margetis, G., Ntoa, S., Antona, M., Stephanidis, C.: Real-time teacher assistance in technologically-augmented smart classrooms Int. J. Adv. Life Sci 6(1), 62\u201373 (2014)","journal-title":"J. Adv. Life Sci"},{"key":"1783_CR6","doi-asserted-by":"crossref","unstructured":"Augmenting physical books towards education enhancement. In: 2013 1st IEEE Workshop on User-Centered Computer Vision (UCCV), pp. 43\u201349 (2013). IEEE","DOI":"10.1109\/UCCV.2013.6530807"},{"key":"1783_CR7","doi-asserted-by":"publisher","first-page":"427","DOI":"10.1007\/s10209-014-0365-0","volume":"14","author":"G Margetis","year":"2015","unstructured":"Margetis, G., Zabulis, X., Ntoa, S., Koutlemanis, P., Papadaki, E., Antona, M., Stephanidis, C.: Enhancing education through natural interaction with physical paper. Univ. Access Inf. Soc. 14, 427\u2013447 (2015)","journal-title":"Univ. Access Inf. Soc."},{"issue":"4","key":"1783_CR8","first-page":"3","volume":"14","author":"A Leonidis","year":"2012","unstructured":"Leonidis, A., Korozi, M., Margetis, G., Ntoa, S., Papagiannakis, H., Antona, M., Stephanidis, C.: A glimpse into the ambient classroom. Bulletin of the IEEE Technical Committee on Learning Technology 14(4), 3 (2012)","journal-title":"Bulletin of the IEEE Technical Committee on Learning Technology"},{"key":"1783_CR9","doi-asserted-by":"publisher","first-page":"87","DOI":"10.12688\/openreseurope.13715.1","volume":"1","author":"KC Apostolakis","year":"2021","unstructured":"Apostolakis, K.C., Dimitriou, N., Margetis, G., Ntoa, S., Tzovaras, D., Stephanidis, C.: Darlene-improving situational awareness of european law enforcement agents through a combination of augmented reality and artificial intelligence solutions. Open Research Europe 1, 87 (2021)","journal-title":"Open Research Europe"},{"key":"1783_CR10","doi-asserted-by":"publisher","first-page":"23367","DOI":"10.1109\/ACCESS.2022.3152743","volume":"10","author":"Z Stefanidi","year":"2022","unstructured":"Stefanidi, Z., Margetis, G., Ntoa, S., Papagiannakis, G.: Real-time adaptation of context-aware intelligent user interfaces, for enhanced situational awareness. IEEE Access 10, 23367\u201323393 (2022). https:\/\/doi.org\/10.1109\/ACCESS.2022.3152743","journal-title":"IEEE Access"},{"issue":"1","key":"1783_CR11","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1007\/s10055-023-00937-2","volume":"28","author":"I Karakostas","year":"2024","unstructured":"Karakostas, I., Valakou, A., Gavgiotaki, D., Stefanidi, Z., Pastaltzidis, I., Tsipouridis, G., Kilis, N., Apostolakis, K.C., Ntoa, S., Dimitriou, N., et al.: A real-time wearable ar system for egocentric vision on the edge. Virtual Reality 28(1), 1\u201324 (2024)","journal-title":"Virtual Reality"},{"key":"1783_CR12","unstructured":"Helin, K., No\u00ebl, F., Sch\u00e4fer, W.: Euroxr 2023: Proceedings of the 20th euroxr international conference. In: 20th EuroXR International Conference (2023). VTT Technical Research Centre of Finland"},{"issue":"10","key":"1783_CR13","doi-asserted-by":"publisher","first-page":"19173","DOI":"10.1109\/TITS.2022.3161141","volume":"23","author":"J Zhang","year":"2022","unstructured":"Zhang, J., Yang, K., Constantinescu, A., Peng, K., M\u00fcller, K., Stiefelhagen, R.: Trans4trans: Efficient transformer for transparent object and semantic scene segmentation in real-world navigation assistance. IEEE Trans. Intell. Transp. Syst. 23(10), 19173\u201319186 (2022)","journal-title":"IEEE Trans. Intell. Transp. Syst."},{"key":"1783_CR14","doi-asserted-by":"publisher","unstructured":"Schieber, H., Kleinbeck, C., Theelke, L., Kraft, M., Kreimeier, J., Roth, D.: Mr-sense: A mixed reality environment search assistant for blind and visually impaired people. In: 2024 IEEE International Conference on Artificial Intelligence and eXtended and Virtual Reality (AIxVR), pp. 166\u2013175 (2024). https:\/\/doi.org\/10.1109\/AIxVR59861.2024.00029","DOI":"10.1109\/AIxVR59861.2024.00029"},{"key":"1783_CR15","doi-asserted-by":"publisher","unstructured":"Mehta, D., Sridhar, S., Sotnychenko, O., Rhodin, H., Shafiei, M., Seidel, H.-P., Xu, W., Casas, D., Theobalt, C.: Vnect: Real-time 3d human pose estimation with a single rgb camera. In: ACM Transactions on Graphics, vol. 36 (2017). https:\/\/doi.org\/10.1145\/3072959.3073596 . http:\/\/gvv.mpi-inf.mpg.de\/projects\/VNect\/","DOI":"10.1145\/3072959.3073596"},{"key":"1783_CR16","doi-asserted-by":"publisher","unstructured":"Mehta, D., Sotnychenko, O., Mueller, F., Xu, W., Elgharib, M., Fua, P., Seidel, H.-P., Rhodin, H., Pons-Moll, G., Theobalt, C.: XNect: Real-time multi-person 3D motion capture with a single RGB camera. In: ACM Transactions on Graphics, vol. 39 (2020). https:\/\/doi.org\/10.1145\/3386569.3392410 . http:\/\/gvv.mpi-inf.mpg.de\/projects\/XNect\/","DOI":"10.1145\/3386569.3392410"},{"key":"1783_CR17","unstructured":"Kun, Z., Xiaoguang, H., Nianjuan, J., Kui, J., Jiangbo, L.: Hemlets pose: Learning part-centric heatmap triplets for accurate 3d human pose estimation. In: Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV) (2019)"},{"issue":"3","key":"1783_CR18","doi-asserted-by":"publisher","first-page":"43","DOI":"10.1007\/s00138-024-01514-6","volume":"35","author":"Y Shi","year":"2024","unstructured":"Shi, Y., Yue, T., Zhao, H., He, G., Ren, K.: Ssman: self-supervised masked adaptive network for 3d human pose estimation. Mach. Vis. Appl. 35(3), 43 (2024)","journal-title":"Mach. Vis. Appl."},{"issue":"6","key":"1783_CR19","doi-asserted-by":"publisher","first-page":"3000","DOI":"10.1109\/TPAMI.2021.3051173","volume":"44","author":"K Zhou","year":"2021","unstructured":"Zhou, K., Han, X., Jiang, N., Jia, K., Lu, J.: Hemlets posh: Learning part-centric heatmap triplets for 3d human pose and shape estimation. IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) 44(6), 3000\u20133014 (2021)","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI)"},{"key":"1783_CR20","doi-asserted-by":"crossref","unstructured":"Rajasegaran, J., Pavlakos, G., Kanazawa, A., Malik, J.: Tracking people by predicting 3D appearance, location & pose. In: CVPR (2022)","DOI":"10.1109\/CVPR52688.2022.00276"},{"key":"1783_CR21","doi-asserted-by":"crossref","unstructured":"Rajasegaran, J., Pavlakos, G., Kanazawa, A., Feichtenhofer, C., Malik, J.: On the benefits of 3d pose and tracking for human action recognition. In: Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp. 640\u2013649 (2023)","DOI":"10.1109\/CVPR52729.2023.00069"},{"key":"1783_CR22","doi-asserted-by":"crossref","unstructured":"Zhang, Y., Ji, P., Kortylewski, A., Wang, A., Mei, J., Yuille, A.L.: 3D-Aware Neural Body Fitting for Occlusion Robust 3D Human Pose Estimation. In: Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV) (2023)","DOI":"10.1109\/ICCV51070.2023.00862"},{"issue":"6","key":"1783_CR23","doi-asserted-by":"publisher","first-page":"82","DOI":"10.1007\/s00138-022-01334-6","volume":"33","author":"H Yang","year":"2022","unstructured":"Yang, H., Guo, L., Zhang, Y., Wu, X.: U-shaped spatial-temporal transformer network for 3d human pose estimation. Mach. Vis. Appl. 33(6), 82 (2022)","journal-title":"Mach. Vis. Appl."},{"key":"1783_CR24","doi-asserted-by":"crossref","unstructured":"Ye, H., Zhu, W., Wang, C., Wu, R., Wang, Y.: Faster voxelpose: Real-time 3d human pose estimation by orthographic projection. In: European Conference on Computer Vision (ECCV) (2022)","DOI":"10.1007\/978-3-031-20068-7_9"},{"key":"1783_CR25","doi-asserted-by":"crossref","unstructured":"Tu, H., Wang, C., Zeng, W.: Voxelpose: Towards multi-camera 3d human pose estimation in wild environment. In: European Conference on Computer Vision (ECCV) (2020)","DOI":"10.1007\/978-3-030-58452-8_12"},{"key":"1783_CR26","doi-asserted-by":"crossref","unstructured":"Iskakov, K., Burkov, E., Lempitsky, V., Malkov, Y.: Learnable triangulation of human pose. In: Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV) (2019)","DOI":"10.1109\/ICCV.2019.00781"},{"issue":"1","key":"1783_CR27","doi-asserted-by":"publisher","first-page":"6","DOI":"10.1007\/s00138-020-01120-2","volume":"32","author":"A Kadkhodamohammadi","year":"2021","unstructured":"Kadkhodamohammadi, A., Padoy, N.: A generalizable approach for multi-view 3d human pose regression. Mach. Vis. Appl. 32(1), 6 (2021)","journal-title":"Mach. Vis. Appl."},{"issue":"3","key":"1783_CR28","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1007\/s00138-024-01530-6","volume":"35","author":"D Rodriguez-Criado","year":"2024","unstructured":"Rodriguez-Criado, D., Bachiller-Burgos, P., Vogiatzis, G., Manso, L.J.: Multi-person 3d pose estimation from unlabelled data. Mach. Vis. Appl. 35(3), 1\u201318 (2024)","journal-title":"Mach. Vis. Appl."},{"key":"1783_CR29","doi-asserted-by":"crossref","unstructured":"Ntoa, S., Margetis, G., Valakou, A., Makri, F., Dimitriou, N., Karakostas, I., Kokkinis, G., Apostolakis, K.C., Tzovaras, D., Stephanidis, C.: A mixed-methods approach for the evaluation of situational awareness and user experience with augmented reality technologies. In: International Conference on Human-Computer Interaction, pp. 199\u2013219 (2024). Springer","DOI":"10.1007\/978-3-031-61569-6_13"},{"key":"1783_CR30","doi-asserted-by":"crossref","unstructured":"Qiu, H., Wang, C., Wang, J., Wang, N., Zeng, W.: Cross view fusion for 3d human pose estimation. In: Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV) (2019)","DOI":"10.1109\/ICCV.2019.00444"},{"key":"1783_CR31","doi-asserted-by":"crossref","unstructured":"Ma, H., Chen, L., Kong, D., Wang, Z., Liu, X., Tang, H., Yan, X., Xie, Y., Lin, S.-Y., Xie, X.: Transfusion: Cross-view fusion with transformer for 3d human pose estimation. In: British Machine Vision Conference (2021)","DOI":"10.5244\/C.35.5"},{"key":"1783_CR32","doi-asserted-by":"crossref","unstructured":"Lin, J., Lee, G.H.: Multi-view multi-person 3d pose estimation with plane sweep stereo. In: Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp. 11886\u201311895 (2021)","DOI":"10.1109\/CVPR46437.2021.01171"},{"issue":"2","key":"1783_CR33","doi-asserted-by":"publisher","first-page":"2613","DOI":"10.1109\/TPAMI.2022.3163709","volume":"45","author":"Y Zhang","year":"2022","unstructured":"Zhang, Y., Wang, C., Wang, X., Liu, W., Zeng, W.: Voxeltrack: Multi-person 3d human pose estimation and tracking in the wild. IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) 45(2), 2613\u20132626 (2022)","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI)"},{"key":"1783_CR34","unstructured":"Wang, T., Zhang, J., Cai, Y., Yan, S., Feng, J.: Direct multi-view multi-person 3d human pose estimation. Advances in Neural Information Processing Systems (2021)"},{"key":"1783_CR35","doi-asserted-by":"crossref","unstructured":"Ruiz, N., Chong, E., Rehg, J.M.: Fine-grained head pose estimation without keypoints. In: The IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Workshops (2018)","DOI":"10.1109\/CVPRW.2018.00281"},{"issue":"4","key":"1783_CR36","doi-asserted-by":"publisher","first-page":"1035","DOI":"10.1109\/TMM.2018.2866770","volume":"21","author":"H-W Hsu","year":"2019","unstructured":"Hsu, H.-W., Wu, T.-Y., Wan, S., Wong, W.H., Lee, C.-Y.: Quatnet: Quaternion-based head pose estimation with multiregression loss. IEEE Trans. Multimedia 21(4), 1035\u20131046 (2019). https:\/\/doi.org\/10.1109\/TMM.2018.2866770","journal-title":"IEEE Trans. Multimedia"},{"key":"1783_CR37","doi-asserted-by":"publisher","DOI":"10.1016\/j.imavis.2019.11.005","volume":"93","author":"B Huang","year":"2020","unstructured":"Huang, B., Chen, R., Xu, W., Zhou, Q.: Improving head pose estimation using two-stage ensembles with top-k regression. Image Vis. Comput. 93, 103827 (2020). https:\/\/doi.org\/10.1016\/j.imavis.2019.11.005","journal-title":"Image Vis. Comput."},{"key":"1783_CR38","doi-asserted-by":"crossref","unstructured":"Zhou, Y., Gregson, J.: WHENet: Real-time Fine-Grained Estimation for Wide Range Head Pose (2020)","DOI":"10.5244\/C.34.189"},{"key":"1783_CR39","doi-asserted-by":"publisher","unstructured":"Yang, T.-Y., Chen, Y.-T., Lin, Y.-Y., Chuang, Y.-Y.: Fsa-net: Learning fine-grained structure aggregation for head pose estimation from a single image. In: 2019 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp. 1087\u20131096 (2019). https:\/\/doi.org\/10.1109\/CVPR.2019.00118","DOI":"10.1109\/CVPR.2019.00118"},{"key":"1783_CR40","doi-asserted-by":"publisher","first-page":"12789","DOI":"10.1609\/aaai.v34i07.6974","volume":"34","author":"H Zhang","year":"2020","unstructured":"Zhang, H., Wang, M., Liu, Y., Yuan, Y.: Fdn: Feature decoupling network for head pose estimation. Proceedings of the AAAI Conference on Artificial Intelligence 34, 12789\u201312796 (2020). https:\/\/doi.org\/10.1609\/aaai.v34i07.6974","journal-title":"Proceedings of the AAAI Conference on Artificial Intelligence"},{"key":"1783_CR41","doi-asserted-by":"crossref","unstructured":"Cao, Z., Chu, Z., Liu, D., Chen, Y.: A vector-based representation to enhance head pose estimation. In: Proceedings of the IEEE\/CVF Winter Conference on Applications of Computer Vision (WACV), pp. 1188\u20131197 (2021)","DOI":"10.1109\/WACV48630.2021.00123"},{"key":"1783_CR42","doi-asserted-by":"crossref","unstructured":"Zhou, Y., Barnes, C., Lu, J., Yang, J., Li, H.: On the continuity of rotation representations in neural networks (2018)","DOI":"10.1109\/CVPR.2019.00589"},{"key":"1783_CR43","doi-asserted-by":"publisher","unstructured":"Hempel, T., Abdelrahman, A.A., Al-Hamadi, A.: 6d rotation representation for unconstrained head pose estimation. In: 2022 IEEE International Conference on Image Processing (ICIP), pp. 2496\u20132500 (2022). https:\/\/doi.org\/10.1109\/ICIP46576.2022.9897219","DOI":"10.1109\/ICIP46576.2022.9897219"},{"key":"1783_CR44","unstructured":"Zhou, H., Jiang, F., Lu, H.: Directmhp: Direct 2d multi-person head pose estimation with full-range angles. arXiv preprint arXiv:2302.01110 (2023)"},{"key":"1783_CR45","doi-asserted-by":"crossref","unstructured":"Olson, E.: Apriltag: A robust and flexible visual fiducial system. In: 2011 IEEE International Conference on Robotics and Automation, pp. 3400\u20133407 (2011). IEEE","DOI":"10.1109\/ICRA.2011.5979561"},{"key":"1783_CR46","doi-asserted-by":"crossref","unstructured":"Wang, J., Olson, E.: Apriltag 2: Efficient and robust fiducial detection. In: 2016 IEEE\/RSJ International Conference on Intelligent Robots and Systems (IROS), pp. 4193\u20134198 (2016). IEEE","DOI":"10.1109\/IROS.2016.7759617"},{"key":"1783_CR47","doi-asserted-by":"crossref","unstructured":"Fiala, M.: Artag, a fiducial marker system using digital techniques. In: 2005 IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR\u201905), vol. 2, pp. 590\u2013596 (2005). IEEE","DOI":"10.1109\/CVPR.2005.74"},{"issue":"6","key":"1783_CR48","doi-asserted-by":"publisher","first-page":"2280","DOI":"10.1016\/j.patcog.2014.01.005","volume":"47","author":"S Garrido-Jurado","year":"2014","unstructured":"Garrido-Jurado, S., Mu\u00f1oz-Salinas, R., Madrid-Cuevas, F.J., Mar\u00edn-Jim\u00e9nez, M.J.: Automatic generation and detection of highly reliable fiducial markers under occlusion. Pattern Recogn. 47(6), 2280\u20132292 (2014)","journal-title":"Pattern Recogn."},{"key":"1783_CR49","doi-asserted-by":"crossref","unstructured":"Kato, H., Billinghurst, M.: Marker tracking and hmd calibration for a video-based augmented reality conferencing system. In: Proceedings 2nd IEEE and ACM International Workshop on Augmented Reality (IWAR\u201999), pp. 85\u201394 (1999). IEEE","DOI":"10.1109\/IWAR.1999.803809"},{"key":"1783_CR50","doi-asserted-by":"crossref","unstructured":"Zhu, Z., Wang, Q., Li, B., Wu, W., Yan, J., Hu, W.: Distractor-aware siamese networks for visual object tracking. In: Proceedings of the European Conference on Computer Vision (ECCV), pp. 101\u2013117 (2018)","DOI":"10.1007\/978-3-030-01240-3_7"},{"key":"1783_CR51","doi-asserted-by":"crossref","unstructured":"Zhang, Y., Sun, P., Jiang, Y., Yu, D., Weng, F., Yuan, Z., Luo, P., Liu, W., Wang, X.: Bytetrack: Multi-object tracking by associating every detection box. In: European Conference on Computer Vision, pp. 1\u201321 (2022). Springer","DOI":"10.1007\/978-3-031-20047-2_1"},{"key":"1783_CR52","unstructured":"Shao, Z., Hoffmann, N., et al: Nam: Normalization-based attention module. In: NeurIPS 2021 Workshop on ImageNet: Past, Present, and Future (2021)"},{"key":"1783_CR53","unstructured":"Joo, H., Simon, T., Li, X., Liu, H., Tan, L., Gui, L., Banerjee, S., Godisart, T.S., Nabbe, B., Matthews, I., Kanade, T., Nobuhara, S., Sheikh, Y.: Panoptic studio: A massively multiview system for social interaction capture. IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) (2017)"},{"key":"1783_CR54","doi-asserted-by":"publisher","unstructured":"Belagiannis, V., Amin, S., Andriluka, M., Schiele, B., Navab, N., Ilic, S.: 3d pictorial structures for multiple human pose estimation. In: Proceedings of the IEEE\/CVF Internatioanl Conference on Computer Vision and Pattern Recognition (CVPR), pp. 1669\u20131676 (2014). https:\/\/doi.org\/10.1109\/CVPR.2014.216","DOI":"10.1109\/CVPR.2014.216"},{"key":"1783_CR55","unstructured":"Xu, Y., Zhang, J., Zhang, Q., Tao, D.: ViTPose: Simple vision transformer baselines for human pose estimation. In: Advances in Neural Information Processing Systems (2022)"},{"key":"1783_CR56","unstructured":"Agarwal, S., Mierle, K., Team, T.C.S.: Ceres Solver. https:\/\/github.com\/ceres-solver\/ceres-solver"},{"key":"1783_CR57","doi-asserted-by":"crossref","unstructured":"Gan, Y., Han, R., Yin, L., Feng, W., Wang, S.: Self-supervised multi-view multi-human association and tracking. In: Proceedings of the 29th ACM International Conference on Multimedia, pp. 282\u2013290 (2021)","DOI":"10.1145\/3474085.3475177"},{"key":"1783_CR58","doi-asserted-by":"crossref","unstructured":"Wojke, N., Bewley, A., Paulus, D.: Simple online and realtime tracking with a deep association metric. In: 2017 IEEE International Conference on Image Processing (ICIP), pp. 3645\u20133649 (2017). IEEE","DOI":"10.1109\/ICIP.2017.8296962"},{"issue":"6","key":"1783_CR59","doi-asserted-by":"publisher","first-page":"381","DOI":"10.1145\/358669.358692","volume":"24","author":"MA Fischler","year":"1981","unstructured":"Fischler, M.A., Bolles, R.C.: Random sample consensus: A paradigm for model fitting with applications to image analysis and automated cartography. Commun. ACM 24(6), 381\u2013395 (1981). https:\/\/doi.org\/10.1145\/358669.358692","journal-title":"Commun. ACM"}],"container-title":["Machine Vision and Applications"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s00138-025-01783-9.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s00138-025-01783-9","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s00138-025-01783-9.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,3,9]],"date-time":"2026-03-09T18:12:48Z","timestamp":1773079968000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s00138-025-01783-9"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,2,6]]},"references-count":59,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2026,3]]}},"alternative-id":["1783"],"URL":"https:\/\/doi.org\/10.1007\/s00138-025-01783-9","relation":{},"ISSN":["0932-8092","1432-1769"],"issn-type":[{"value":"0932-8092","type":"print"},{"value":"1432-1769","type":"electronic"}],"subject":[],"published":{"date-parts":[[2026,2,6]]},"assertion":[{"value":"7 January 2025","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"9 December 2025","order":2,"name":"revised","label":"Revised","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"24 December 2025","order":3,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"6 February 2026","order":4,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"All authors have no Conflict of interest.","order":2,"name":"Ethics","group":{"name":"EthicsHeading","label":"Conflict of interest"}}],"article-number":"31"}}