{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,23]],"date-time":"2026-04-23T18:29:36Z","timestamp":1776968976083,"version":"3.51.4"},"reference-count":30,"publisher":"MDPI AG","issue":"9","license":[{"start":{"date-parts":[[2015,9,22]],"date-time":"2015-09-22T00:00:00Z","timestamp":1442880000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Sensors"],"abstract":"<jats:p>Accurate motion capture plays an important role in sports analysis, the medical field and virtual reality. Current methods for motion capture often suffer from occlusions, which limits the accuracy of their pose estimation. In this paper, we propose a complete system to measure the pose parameters of the human body accurately. Different from previous monocular depth camera systems, we leverage two Kinect sensors to acquire more information about human movements, which ensures that we can still get an accurate estimation even when significant occlusion occurs. Because human motion is temporally constant, we adopt a learning analysis to mine the temporal information across the posture variations. Using this information, we estimate human pose parameters accurately, regardless of rapid movement. Our experimental results show that our system can perform an accurate pose estimation of the human body with the constraint of information from the temporal domain.<\/jats:p>","DOI":"10.3390\/s150924297","type":"journal-article","created":{"date-parts":[[2015,9,22]],"date-time":"2015-09-22T15:36:54Z","timestamp":1442936214000},"page":"24297-24317","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":35,"title":["Leveraging Two Kinect Sensors for Accurate Full-Body Motion Capture"],"prefix":"10.3390","volume":"15","author":[{"given":"Zhiquan","family":"Gao","sequence":"first","affiliation":[{"name":"School of Electronic Science and Engineering, Nanjing University, Nanjing 210046, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yao","family":"Yu","sequence":"additional","affiliation":[{"name":"School of Electronic Science and Engineering, Nanjing University, Nanjing 210046, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yu","family":"Zhou","sequence":"additional","affiliation":[{"name":"School of Electronic Science and Engineering, Nanjing University, Nanjing 210046, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Sidan","family":"Du","sequence":"additional","affiliation":[{"name":"School of Electronic Science and Engineering, Nanjing University, Nanjing 210046, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1968","published-online":{"date-parts":[[2015,9,22]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","first-page":"90","DOI":"10.1016\/j.cviu.2006.08.002","article-title":"A Survey of Advances in Vision-Based Human Motion Capture and Analysis","volume":"104","author":"Moeslund","year":"2006","journal-title":"Comp. Vis. Image Uud."},{"key":"ref_2","unstructured":"Vicon System 2014. Available online: http:\/\/www.vicon.com\/."},{"key":"ref_3","unstructured":"Xsens 2014. Available online: http:\/\/www.xsens.com\/."},{"key":"ref_4","unstructured":"Ascension 2014. Available online: http:\/\/www.ascension-tech.com\/."},{"key":"ref_5","first-page":"1","article-title":"Performance Capture from Sparse Multi-view Video","volume":"27","author":"Stoll","year":"2008","journal-title":"ACM Trans. Graph."},{"key":"ref_6","doi-asserted-by":"crossref","unstructured":"Liu, Y., Stoll, C., Gall, J., Seidel, H.P., and Theobalt, C. (2011, January 20\u201325). Markerless Motion Capture of Interacting Characters using Multi-View Image Segmentation. Proceedings of the IEEE International Conference on Computer Vision and Pattern Recognition, Colorado Springs, CO, USA.","DOI":"10.1109\/CVPR.2011.5995424"},{"key":"ref_7","doi-asserted-by":"crossref","unstructured":"Stoll, C., Hasler, N., Gall, J., Seidel, H., and Theobalt, C. (2011, January 6\u201313). Fast Articulated Motion Tracking using a Sums of Gaussians Body Model. Proceedings of the IEEE International Conference on Computer Vision, Barcelona, Spain.","DOI":"10.1109\/ICCV.2011.6126338"},{"key":"ref_8","doi-asserted-by":"crossref","unstructured":"Straka, M., Hauswiesner, S., Ruther, M., and Bischof, H. (2012, January 13\u201315). Rapid Skin: Estimating the 3D Human Pose and Shape in Real-Time. Proceedings of the IEEE International Conference on 3D Imaging, Modeling, Processing, Visualization and Transmission, ETH Z\u00fcrich, Switzerland.","DOI":"10.1109\/3DIMPVT.2012.18"},{"key":"ref_9","doi-asserted-by":"crossref","first-page":"1437","DOI":"10.3390\/s120201437","article-title":"Accuracy and Resolution of Kinect Depth Data for Indoor Mapping Applications","volume":"12","author":"Khoshelham","year":"2012","journal-title":"Sensors"},{"key":"ref_10","doi-asserted-by":"crossref","unstructured":"Ye, M., Wang, X., Yang, R., Ren, L., and Pollefeys, M. (2011, January 6\u201313). Accurate 3D Pose Estimation from a Single Depth Image. Proceedings of the IEEE International Conference on Computer Vision, Barcelona, Spain.","DOI":"10.1109\/ICCV.2011.6126310"},{"key":"ref_11","doi-asserted-by":"crossref","unstructured":"Weiss, A., Hirshberg, D., and Black, M.J. (2011, January 6\u201313). Home 3D Body Scans from Noisy Image and Range Data. Proceedings of the IEEE International Conference on Computer Vision, Barcelona, Spain.","DOI":"10.1109\/ICCV.2011.6126465"},{"key":"ref_12","doi-asserted-by":"crossref","first-page":"116","DOI":"10.1145\/2398356.2398381","article-title":"Real-Time Human Pose Recognition in Parts from Single Depth Images","volume":"56","author":"Shotton","year":"2013","journal-title":"Commun. ACM"},{"key":"ref_13","unstructured":"Grest, D., Kr\u00fcger, V., and Koch, R. (2007). Image Analysis, Springer."},{"key":"ref_14","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/2366145.2366207","article-title":"Accurate Realtime Full-Body Motion Capture using a Single Depth Camera","volume":"31","author":"Wei","year":"2012","journal-title":"ACM Trans. Graph."},{"key":"ref_15","doi-asserted-by":"crossref","first-page":"11362","DOI":"10.3390\/s130911362","article-title":"Measuring Accurate Body Parameters of Dressed Humans with Large-Scale Motion Using a Kinect Sensor","volume":"13","author":"Xu","year":"2013","journal-title":"Sensors"},{"key":"ref_16","doi-asserted-by":"crossref","first-page":"408","DOI":"10.1145\/1073204.1073207","article-title":"SCAPE: Shape Completion and Animation of PEople","volume":"24","author":"Anguelov","year":"2005","journal-title":"ACM Trans. Graph."},{"key":"ref_17","doi-asserted-by":"crossref","unstructured":"Shen, W., Deng, K., Bai, X., Leyvand, T., Guo, B., and Tu, Z. (2012, January 16\u201321). Exemplar-based human action pose correction and tagging. Proceedings of the IEEE International Conference on Computer Vision and Pattern Recognition (CVPR), Providence, RI, USA.","DOI":"10.1109\/CVPR.2012.6247875"},{"key":"ref_18","doi-asserted-by":"crossref","first-page":"1053","DOI":"10.1109\/TCYB.2013.2279071","article-title":"Exemplar-based human action pose correction","volume":"44","author":"Shen","year":"2014","journal-title":"IEEE Trans. Cybern."},{"key":"ref_19","unstructured":"Shen, W., Lei, R., Zeng, D., and Zhang, Z. (2014). Computer Vision\u2014ACCV 2014, Springer."},{"key":"ref_20","unstructured":"Microsoft Kinect API for Windows. Available online: https:\/\/www.microsoft.com\/en-us\/kinectforwindows\/."},{"key":"ref_21","doi-asserted-by":"crossref","unstructured":"Essmaeel, K., Gallo, L., Damiani, E., de Pietro, G., and Dipanda, A. (2012, January 25\u201329). Temporal denoising of kinect depth data. Proceedings of the 8th IEEE International Conference on Signal Image Technology and Internet Based Systems (SITIS), Naples, Italy.","DOI":"10.1109\/SITIS.2012.18"},{"key":"ref_22","doi-asserted-by":"crossref","first-page":"587","DOI":"10.1145\/882262.882311","article-title":"The space of human body shapes: Reconstruction and parameterization from range scans","volume":"22","author":"Allen","year":"2003","journal-title":"ACM Trans. Graph. (TOG)"},{"key":"ref_23","doi-asserted-by":"crossref","unstructured":"Yang, Y., Yu, Y., Zhou, Y., Sidan, D., Davis, J., and Yang, R. (2014, January 8\u201311). Semantic Parametric Reshaping of Human Body Models. Proceedings of the 2014 2nd International Conference on 3D Vision, Tokyo, Japan.","DOI":"10.1109\/3DV.2014.47"},{"key":"ref_24","doi-asserted-by":"crossref","unstructured":"Hasler, N., Stoll, C., Sunkel, M., Rosenhahn, B., and Seidel, H.P. (2009, January 3). A Statistical Model of Human Pose and Body Shape. Proceedings of the Annual Conference of the European Association for Computer Graphics, Munich, Germany.","DOI":"10.1111\/j.1467-8659.2009.01373.x"},{"key":"ref_25","doi-asserted-by":"crossref","unstructured":"Desbrun, M., Meyer, M., Schr\u00f6der, P., and Barr, A.H. (1999, January 8\u201313). Implicit Fairing of Irregular Meshes using Diffusion and Curvature Flow. Proceedings of the 26th Annual Conference on Computer Graphics and Interactive Techniques, Los Angeles, CA, USA.","DOI":"10.1145\/311535.311576"},{"key":"ref_26","unstructured":"Berger, K., Ruhl, K., Schroeder, Y., Bruemmer, C., Scholz, A., and Magnor, M.A. (2011, January 4\u20136). Markerless Motion Capture Using Multiple Color-Depth Sensors. Proceedings of the International Workshop on Vision, Modeling, and Visualization, Berlin, German."},{"key":"ref_27","doi-asserted-by":"crossref","unstructured":"Auvinet, E., Meunier, J., and Multon, F. (2012, January 2\u20135). Multiple depth cameras calibration and body volume reconstruction for gait analysis. Proceedings of the IEEE 2012 11th International Conference on Information Science, Signal Processing and their Applications (ISSPA), Montreal, QC, Canada.","DOI":"10.1109\/ISSPA.2012.6310598"},{"key":"ref_28","doi-asserted-by":"crossref","first-page":"291","DOI":"10.1109\/89.279278","article-title":"Maximum a Posteriori Estimation for Multivariate Gaussian Mixture Observations of Markov chains","volume":"2","author":"Gauvain","year":"1994","journal-title":"IEEE Trans. Audio Speech Lang. Process."},{"key":"ref_29","doi-asserted-by":"crossref","first-page":"2262","DOI":"10.1109\/TPAMI.2010.46","article-title":"Point Set Registration: Coherent Point Drift","volume":"32","author":"Myronenko","year":"2010","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_30","doi-asserted-by":"crossref","first-page":"221","DOI":"10.1023\/B:VISI.0000011205.11775.fd","article-title":"Lucas-Kanade 20 Years on: A Unifying Framework","volume":"56","author":"Baker","year":"2004","journal-title":"Int. J. Comp. Vis."}],"container-title":["Sensors"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/1424-8220\/15\/9\/24297\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,11]],"date-time":"2025-10-11T20:49:00Z","timestamp":1760215740000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/1424-8220\/15\/9\/24297"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2015,9,22]]},"references-count":30,"journal-issue":{"issue":"9","published-online":{"date-parts":[[2015,9]]}},"alternative-id":["s150924297"],"URL":"https:\/\/doi.org\/10.3390\/s150924297","relation":{},"ISSN":["1424-8220"],"issn-type":[{"value":"1424-8220","type":"electronic"}],"subject":[],"published":{"date-parts":[[2015,9,22]]}}}