{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,7]],"date-time":"2026-04-07T20:37:56Z","timestamp":1775594276907,"version":"3.50.1"},"reference-count":53,"publisher":"Frontiers Media SA","license":[{"start":{"date-parts":[[2025,3,5]],"date-time":"2025-03-05T00:00:00Z","timestamp":1741132800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"content-domain":{"domain":["frontiersin.org"],"crossmark-restriction":true},"short-container-title":["Front. Neurorobot."],"abstract":"<jats:sec><jats:title>Objective<\/jats:title><jats:p>To address the limitations of traditional methods in human pose recognition, such as occlusions, lighting variations, and motion continuity, particularly in complex dynamic environments for seamless human-robot interaction.<\/jats:p><\/jats:sec><jats:sec><jats:title>Method<\/jats:title><jats:p>We propose PoseRL-Net, a deep learning-based pose recognition model that enhances accuracy and robustness in human pose estimation. PoseRL-Net integrates multiple components, including a Spatial-Temporal Graph Convolutional Network (STGCN), attention mechanism, Gated Recurrent Unit (GRU) module, pose refinement, and symmetry constraints. The STGCN extracts spatial and temporal features, the attention mechanism focuses on key pose features, the GRU ensures temporal consistency, and the refinement and symmetry constraints improve structural plausibility and stability.<\/jats:p><\/jats:sec><jats:sec><jats:title>Results<\/jats:title><jats:p>Extensive experiments conducted on the Human3.6M and MPI-INF-3DHP datasets demonstrate that PoseRL-Net outperforms existing state-of-the-art models on key metrics such as MPIPE and P-MPIPE, showcasing superior performance across various pose recognition tasks.<\/jats:p><\/jats:sec><jats:sec><jats:title>Conclusion<\/jats:title><jats:p>PoseRL-Net not only improves pose estimation accuracy but also provides crucial support for intelligent decision-making and motion planning in robots operating in dynamic and complex scenarios, offering significant practical value for collaborative robotics.<\/jats:p><\/jats:sec>","DOI":"10.3389\/fnbot.2025.1531894","type":"journal-article","created":{"date-parts":[[2025,3,5]],"date-time":"2025-03-05T07:12:00Z","timestamp":1741158720000},"update-policy":"https:\/\/doi.org\/10.3389\/crossmark-policy","source":"Crossref","is-referenced-by-count":3,"title":["PoseRL-Net: human pose analysis for motion training guided by robot vision"],"prefix":"10.3389","volume":"19","author":[{"given":"Bin","family":"Liu","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Hui","family":"Wang","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1965","published-online":{"date-parts":[[2025,3,5]]},"reference":[{"key":"B1","doi-asserted-by":"publisher","first-page":"104753","DOI":"10.1016\/j.robot.2024.104753","article-title":"Depth accuracy analysis of the ZED 2I stereo camera in an indoor environment","volume":"179","author":"Abdelsalam","year":"2024","journal-title":"Rob. Auton. Syst"},{"key":"B2","doi-asserted-by":"crossref","first-page":"532","DOI":"10.1109\/RO-MAN53752.2022.9900548","article-title":"\u201cFusion of depth, color, and thermal images towards digital twins and safe human interaction with a robot in an industrial environment,\u201d","volume-title":"2022 31st IEEE International Conference on Robot and Human Interactive Communication (RO-MAN)","author":"Al Naser","year":"2022"},{"key":"B3","doi-asserted-by":"publisher","first-page":"11152","DOI":"10.1109\/TPAMI.2024.3458921","article-title":"Learning from human attention for attribute-assisted visual recognition","volume":"46","author":"Bai","year":"2024","journal-title":"IEEE Trans. Pattern. Anal. Mach. Intell"},{"key":"B4","doi-asserted-by":"publisher","first-page":"18839","DOI":"10.1007\/s11042-021-10646-0","article-title":"2D object recognition: a comparative analysis of sift, surf and orb feature descriptors","volume":"80","author":"Bansal","year":"2021","journal-title":"Multimed. Tools Appl"},{"key":"B5","doi-asserted-by":"publisher","first-page":"130","DOI":"10.1016\/j.cviu.2012.10.008","article-title":"Symmetry-driven accumulation of local features for human characterization and re-identification","volume":"117","author":"Bazzani","year":"2013","journal-title":"Comput. Vis. Image Understand"},{"key":"B6","doi-asserted-by":"publisher","first-page":"7291","DOI":"10.1109\/CVPR.2017.143","article-title":"\u201cRealtime multi-person 2D pose estimation using part affinity fields,\u201d","author":"Cao","year":"2017","journal-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition"},{"key":"B7","first-page":"886","article-title":"\u201cHistograms of oriented gradients for human detection,\u201d","volume-title":"2005 IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR'05)","author":"Dalal","year":"2005"},{"key":"B8","doi-asserted-by":"publisher","first-page":"17064","DOI":"10.1109\/JSEN.2021.3081188","article-title":"Validation of a 3D markerless system for gait analysis based on openPose and two RGB webcams","volume":"21","author":"D'Antonio","year":"2021","journal-title":"IEEE Sens. J"},{"key":"B9","doi-asserted-by":"publisher","first-page":"103275","DOI":"10.1016\/j.cviu.2021.103275","article-title":"A review of 3D human pose estimation algorithms for markerless motion capture","volume":"212","author":"Desmarais","year":"2021","journal-title":"Comput. Vis. Image Underst"},{"key":"B10","doi-asserted-by":"publisher","first-page":"16940","DOI":"10.1109\/TITS.2022.3160741","article-title":"Towards real-time monocular depth estimation for robotics: a survey","volume":"23","author":"Dong","year":"2022","journal-title":"IEEE Trans. Intell. Transp. Syst"},{"key":"B11","doi-asserted-by":"publisher","first-page":"102304","DOI":"10.1016\/j.rcim.2021.102304","article-title":"Vision-based holistic scene understanding towards proactive human-robot collaboration","volume":"75","author":"Fan","year":"2022","journal-title":"Robot. Comput. Integr. Manuf"},{"key":"B12","doi-asserted-by":"publisher","first-page":"7157","DOI":"10.1109\/TPAMI.2022.3222784","article-title":"Alphapose: whole-body regional multi-person pose estimation and tracking in real-time","volume":"45","author":"Fang","year":"2022","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell"},{"key":"B13","doi-asserted-by":"publisher","first-page":"102551","DOI":"10.1016\/j.inffus.2024.102551","article-title":"Coarse to fine-based image\u2013point cloud fusion network for 3D object detection","volume":"112","author":"Hao","year":"2024","journal-title":"Inf. Fusion"},{"key":"B14","doi-asserted-by":"publisher","first-page":"637","DOI":"10.5220\/0012383700003657","article-title":"\u201cGait parameter estimation from a single privacy preserving depth sensor,\u201d","author":"Hartmann","year":"2024","journal-title":"BIOSTEC"},{"key":"B15","doi-asserted-by":"publisher","first-page":"40","DOI":"10.5220\/0010840500003123","article-title":"\u201cInterpretable high-level features for human activity recognition,\u201d","author":"Hartmann","year":"2022","journal-title":"Biosignals"},{"key":"B16","doi-asserted-by":"publisher","first-page":"135","DOI":"10.5220\/0008851401350140","article-title":"\u201cFeature space reduction for multimodal human activity recognition,\u201d","author":"Hartmann","year":"2020","journal-title":"Proceedings of the 13th International Joint Conference on Biomedical Engineering Systems and Technologies (BIOSTEC 2020)"},{"key":"B17","doi-asserted-by":"publisher","first-page":"215","DOI":"10.5220\/0010260800002865","article-title":"\u201cFeature space reduction for human activity recognition based on multi-channel biosignals,\u201d","author":"Hartmann","year":"2021","journal-title":"BIOSIGNALS"},{"key":"B18","doi-asserted-by":"crossref","first-page":"141","DOI":"10.1007\/978-3-031-38854-5_8","article-title":"\u201cHigh-level features for human activity recognition and modeling,\u201d","volume-title":"Biomedical Engineering Systems and Technologies","author":"Hartmann","year":"2023"},{"key":"B19","doi-asserted-by":"publisher","first-page":"4212","DOI":"10.1109\/TIP.2023.3275914","article-title":"Regular splitting graph network for 3D human pose estimation","volume":"32","author":"Hassan","year":"2023","journal-title":"IEEE Trans. Image Proc"},{"key":"B20","doi-asserted-by":"publisher","first-page":"770","DOI":"10.1109\/CVPR.2016.90","article-title":"\u201cDeep residual learning for image recognition,\u201d","author":"He","year":"2016","journal-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition"},{"key":"B21","doi-asserted-by":"publisher","first-page":"128913","DOI":"10.1016\/j.physa.2023.128913","article-title":"STGC-GNNS: a GNN-based traffic prediction framework with a spatial-temporal granger causality graph","volume":"623","author":"He","year":"2023","journal-title":"Phys. A"},{"key":"B22","doi-asserted-by":"publisher","first-page":"4183","DOI":"10.3390\/app11094183","article-title":"Human pose detection for robotic-assisted and rehabilitation environments","volume":"11","author":"Hern\u00e1ndez","year":"2021","journal-title":"Appl. Sci"},{"key":"B23","doi-asserted-by":"publisher","first-page":"1325","DOI":"10.1109\/TPAMI.2013.248","article-title":"Human3.6m: large scale datasets and predictive methods for 3d human sensing in natural environments","volume":"36","author":"Ionescu","year":"2013","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell"},{"key":"B24","doi-asserted-by":"publisher","first-page":"2938","DOI":"10.1109\/ICCV.2015.336","article-title":"\u201cPosenet: a convolutional network for real-time 6-dof camera relocalization,\u201d","author":"Kendall","year":"2015","journal-title":"Proceedings of the IEEE International Conference on Computer Vision"},{"key":"B25","doi-asserted-by":"publisher","first-page":"2700","DOI":"10.3390\/app13042700","article-title":"Human pose estimation using mediapipe pose and optimization method based on a humanoid model","volume":"13","author":"Kim","year":"2023","journal-title":"Appl. Sci"},{"key":"B26","doi-asserted-by":"publisher","first-page":"11127","DOI":"10.1109\/ICCV48922.2021.01094","article-title":"\u201cPare: part attention regressor for 3D human body estimation,\u201d","author":"Kocabas","year":"2021","journal-title":"Proceedings of the IEEE\/CVF International Conference on Computer Vision"},{"key":"B27","doi-asserted-by":"publisher","first-page":"2886","DOI":"10.3390\/s20102886","article-title":"Real-time human action recognition with a low-cost RGB camera and mobile robot platform","volume":"20","author":"Lee","year":"2020","journal-title":"Sensors"},{"key":"B28","doi-asserted-by":"publisher","first-page":"1233341","DOI":"10.3389\/fphys.2023.1233341","article-title":"Outlier detection using iterative adaptive mini-minimum spanning tree generation with applications on medical data","volume":"14","author":"Li","year":"2023","journal-title":"Front. Physiol"},{"key":"B29","doi-asserted-by":"publisher","first-page":"015025","DOI":"10.1088\/2632-2153\/ad2492","article-title":"MS2OD: outlier detection using minimum spanning tree and medoid selection","volume":"5","author":"Li","year":"2024","journal-title":"Mach. Learn"},{"key":"B30","doi-asserted-by":"publisher","first-page":"3316","DOI":"10.1109\/TPAMI.2021.3053765","article-title":"Symbiotic graph neural networks for 3d skeleton-based human action recognition and motion prediction","volume":"44","author":"Li","year":"2021","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell"},{"key":"B31","doi-asserted-by":"publisher","first-page":"7107","DOI":"10.1109\/TII.2022.3143605","article-title":"ARHPE: asymmetric relation-aware representation learning for head pose estimation in industrial human-computer interaction","volume":"18","author":"Liu","year":"2022","journal-title":"IEEE Trans. Ind. Inf"},{"key":"B32","doi-asserted-by":"publisher","first-page":"101997","DOI":"10.1016\/j.rcim.2020.101997","article-title":"Collision-free human-robot collaboration based on context awareness","volume":"67","author":"Liu","year":"2021","journal-title":"Robot. Comput. Integr. Manuf"},{"key":"B33","doi-asserted-by":"crossref","first-page":"1150","DOI":"10.1109\/ICCV.1999.790410","article-title":"\u201cObject recognition from local scale-invariant features,\u201d","volume-title":"Proceedings of the Seventh IEEE International Conference on Computer Vision","author":"Lowe","year":"1999"},{"key":"B34","doi-asserted-by":"publisher","first-page":"119580","DOI":"10.1016\/j.ins.2023.119580","article-title":"HISTGNN: hierarchical spatio-temporal graph neural network for weather forecasting","volume":"648","author":"Ma","year":"2023","journal-title":"Inf. Sci"},{"key":"B35","doi-asserted-by":"publisher","first-page":"2640","DOI":"10.1109\/ICCV.2017.288","author":"Martinez","year":"2017","journal-title":"Proceedings of the IEEE International Conference on Computer Vision"},{"key":"B36","doi-asserted-by":"crossref","first-page":"506","DOI":"10.1109\/3DV.2017.00064","article-title":"\u201cMonocular 3D human pose estimation in the wild using improved CNN supervision,\u201d","volume-title":"2017 international conference on 3D vision (3DV)","author":"Mehta","year":"2017"},{"key":"B37","doi-asserted-by":"crossref","first-page":"362","DOI":"10.1109\/ICAIIC48513.2020.9065078","article-title":"\u201cA CNN-LSTM approach to human activity recognition,\u201d","volume-title":"2020 International Conference on Artificial Intelligence in Information and Communication (ICAIIC)","author":"Mutegeki","year":"2020"},{"key":"B38","first-page":"268","article-title":"AI-driven decision support systems in management: enhancing strategic planning and execution","volume":"12","author":"Narneg","year":"2024","journal-title":"Int. J. Rec. Innov. Trends Comput. Commun"},{"key":"B39","doi-asserted-by":"publisher","first-page":"102033","DOI":"10.1016\/j.inffus.2023.102033","article-title":"DILF: differentiable rendering-based multi-view image\u2013language fusion for zero-shot 3d shape understanding","volume":"102","author":"Ning","year":"2024","journal-title":"Inf. Fusion"},{"key":"B40","doi-asserted-by":"publisher","first-page":"7753","DOI":"10.1109\/CVPR.2019.00794","article-title":"\u201c3D human pose estimation in video with temporal convolutions and semi-supervised training,\u201d","author":"Pavllo","year":"2019","journal-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition"},{"key":"B41","doi-asserted-by":"publisher","first-page":"1473","DOI":"10.1109\/TCSVT.2008.2005594","article-title":"Machine recognition of human activities: a survey","volume":"18","author":"Turaga","year":"2008","journal-title":"IEEE Trans. Circ. Syst. Video Technol"},{"key":"B42","article-title":"Graph attention networks","author":"Veli\u010dkovi\u0107","year":"2017","journal-title":"arXiv preprint arXiv:1710.10903"},{"key":"B43","doi-asserted-by":"publisher","first-page":"623","DOI":"10.1016\/j.procs.2015.04.095","article-title":"Cloud based big data analytics framework for face recognition in social networks using machine learning","volume":"50","author":"Vinay","year":"2015","journal-title":"Procedia Comput. Sci"},{"key":"B44","doi-asserted-by":"publisher","first-page":"103225","DOI":"10.1016\/j.cviu.2021.103225","article-title":"Deep 3D human pose estimation: a review","volume":"210","author":"Wang","year":"2021","journal-title":"Comput. Vis. Image Underst"},{"key":"B45","doi-asserted-by":"publisher","first-page":"304","DOI":"10.1109\/THMS.2017.2776211","article-title":"Recognition and detection of two-person interactive actions using automatically selected skeleton features","volume":"48","author":"Wu","year":"2017","journal-title":"IEEE Trans. Hum. Mach. Syst"},{"key":"B46","doi-asserted-by":"publisher","first-page":"16105","DOI":"10.1109\/CVPR46437.2021.01584","article-title":"\u201cGraph stacked hourglass networks for 3D human pose estimation,\u201d","author":"Xu","year":"2021","journal-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition"},{"key":"B47","article-title":"Spatio-temporal graph convolutional networks: a deep learning framework for traffic forecasting","author":"Yu","year":"2017","journal-title":"arXiv preprint arXiv:1709.04875"},{"key":"B48","doi-asserted-by":"publisher","first-page":"1354","DOI":"10.3390\/s19061354","article-title":"Smart sensing and adaptive reasoning for enabling industrial robots with interactive human-robot capabilities in dynamic environments\u2013a case study","volume":"19","author":"Zabalza","year":"2019","journal-title":"Sensors"},{"key":"B49","doi-asserted-by":"publisher","first-page":"9","DOI":"10.1016\/j.cirp.2020.04.077","article-title":"Recurrent neural network for motion trajectory prediction in human-robot collaborative assembly","volume":"69","author":"Zhang","year":"2020","journal-title":"CIRP Ann"},{"key":"B50","doi-asserted-by":"publisher","first-page":"6617286","DOI":"10.1155\/2021\/6617286","article-title":"A 3D machine vision-enabled intelligent robot architecture","volume":"2021","author":"Zhang","year":"2021","journal-title":"Mobile Inf. Syst"},{"key":"B51","doi-asserted-by":"publisher","first-page":"3425","DOI":"10.1109\/CVPR.2019.00354","article-title":"\u201cSemantic graph convolutional networks for 3D human pose regression,\u201d","author":"Zhao","year":"2019","journal-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition"},{"key":"B52","doi-asserted-by":"publisher","first-page":"33040","DOI":"10.1109\/JIOT.2024.3420789","article-title":"MSS-former: multi-scale skeletal transformer for intelligent fall risk prediction in older adults","volume":"11","author":"Zhao","year":"2024","journal-title":"IEEE Internet Things J"},{"key":"B53","doi-asserted-by":"publisher","first-page":"11477","DOI":"10.1109\/ICCV48922.2021.01128","article-title":"\u201cModulated graph convolutional network for 3D human pose estimation,\u201d","author":"Zou","year":"2021","journal-title":"Proceedings of the IEEE\/CVF International Conference on Computer Vision"}],"container-title":["Frontiers in Neurorobotics"],"original-title":[],"link":[{"URL":"https:\/\/www.frontiersin.org\/articles\/10.3389\/fnbot.2025.1531894\/full","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,3,5]],"date-time":"2025-03-05T07:12:06Z","timestamp":1741158726000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.frontiersin.org\/articles\/10.3389\/fnbot.2025.1531894\/full"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,3,5]]},"references-count":53,"alternative-id":["10.3389\/fnbot.2025.1531894"],"URL":"https:\/\/doi.org\/10.3389\/fnbot.2025.1531894","relation":{},"ISSN":["1662-5218"],"issn-type":[{"value":"1662-5218","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,3,5]]},"article-number":"1531894"}}