{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,3,25]],"date-time":"2026-03-25T23:07:06Z","timestamp":1774480026635,"version":"3.50.1"},"reference-count":39,"publisher":"Frontiers Media SA","license":[{"start":{"date-parts":[[2023,6,13]],"date-time":"2023-06-13T00:00:00Z","timestamp":1686614400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"content-domain":{"domain":["frontiersin.org"],"crossmark-restriction":true},"short-container-title":["Front. Neurorobot."],"abstract":"<jats:p>Accurately estimating the 6DoF pose of objects during robot grasping is a common problem in robotics. However, the accuracy of the estimated pose can be compromised during or after grasping the object when the gripper collides with other parts or occludes the view. Many approaches to improving pose estimation involve using multi-view methods that capture RGB images from multiple cameras and fuse the data. While effective, these methods can be complex and costly to implement. In this paper, we present a Single-Camera Multi-View (SCMV) method that utilizes just one fixed monocular camera and the initiative motion of robotic manipulator to capture multi-view RGB image sequences. Our method achieves more accurate 6DoF pose estimation results. We further create a new T-LESS-GRASP-MV dataset specifically for validating the robustness of our approach. Experiments show that the proposed approach outperforms many other public algorithms by a large margin. Quantitative experiments on a real robot manipulator demonstrate the high pose estimation accuracy of our method. Finally, the robustness of the proposed approach is demonstrated by successfully completing an assembly task on a real robot platform, achieving an assembly success rate of 80%.<\/jats:p>","DOI":"10.3389\/fnbot.2023.1136882","type":"journal-article","created":{"date-parts":[[2023,6,13]],"date-time":"2023-06-13T04:13:44Z","timestamp":1686629624000},"update-policy":"https:\/\/doi.org\/10.3389\/crossmark-policy","source":"Crossref","is-referenced-by-count":9,"title":["Single-Camera Multi-View 6DoF pose estimation for robotic grasping"],"prefix":"10.3389","volume":"17","author":[{"given":"Shuangjie","family":"Yuan","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Zhenpeng","family":"Ge","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Lu","family":"Yang","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1965","published-online":{"date-parts":[[2023,6,13]]},"reference":[{"key":"B1","doi-asserted-by":"publisher","first-page":"196","DOI":"10.1504\/IJTM.2015.068224","article-title":"Strategic business transformation through technology convergence: implications from general electric's industrial internet initiative","volume":"67","author":"Agarwal","year":"2015","journal-title":"Int. J. Technol. Manage."},{"key":"B2","doi-asserted-by":"crossref","first-page":"585","DOI":"10.1109\/ICCVW.2011.6130296","article-title":"\u201cCAD-model recognition and 6DoF pose estimation using 3D cues,\u201d","volume-title":"2011 IEEE International Conference on Computer Vision Workshops (ICCV Workshops)","author":"Aldoma","year":"2011"},{"key":"B3","first-page":"7163","article-title":"\u201cPointNetLK: robust & efficient point cloud registration using pointNet,\u201d","author":"Aoki","year":"2019","journal-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition"},{"key":"B4","first-page":"404","article-title":"\u201cSurf: Speeded up robust features,\u201d","volume-title":"Computer Vision - ECCV 2006. ECCV 2006. Lecture Notes in Computer Science, Vol 3951","author":"Bay","year":"2006"},{"key":"B5","doi-asserted-by":"publisher","DOI":"10.48550\/arXiv.1802.10367","article-title":"Deep-6DPose: recovering 6D object pose from a single RGB image","author":"Do","year":"2018","journal-title":"arXiv preprint arXiv:1802.10367"},{"key":"B6","doi-asserted-by":"crossref","first-page":"381","DOI":"10.1145\/358669.358692","article-title":"Random sample consensus: a paradigm for model fitting with applications to image analysis and automated cartography","volume":"24","author":"Fischler","year":"1981","journal-title":"Commun. ACM"},{"key":"B7","first-page":"224","article-title":"\u201cRecognizing objects in range data using regional point descriptors,\u201d","volume-title":"European Conference on Computer Vision","author":"Frome","year":"2004"},{"key":"B8","first-page":"548","article-title":"\u201cModel based training, detection and pose estimation of texture-less 3D objects in heavily cluttered scenes,\u201d","volume-title":"Asian Conference on Computer Vision","author":"Hinterstoisser","year":"2012"},{"key":"B9","doi-asserted-by":"crossref","first-page":"880","DOI":"10.1109\/WACV.2017.103","article-title":"\u201cT-LESS: an RGB-D dataset for 6D pose estimation of texture-less objects,\u201d","volume-title":"2017 IEEE Winter Conference on Applications of Computer Vision (WACV)","author":"Hodan","year":"2017"},{"key":"B10","volume-title":"Spin-Images: A Representation for 3-D Surface Matching","author":"Johnson","year":"1997"},{"key":"B11","first-page":"1521","article-title":"\u201cSSD-6D:making rgb-based 3D detection and 6D pose estimation great again,\u201d","author":"Kehl","year":"2017","journal-title":"Proceedings of the IEEE International Conference on Computer Vision"},{"key":"B12","doi-asserted-by":"crossref","first-page":"3607","DOI":"10.1109\/ICRA.2011.5979949","article-title":"\u201cG 2 o: a general framework for graph optimization,\u201d","volume-title":"2011 IEEE International Conference on Robotics and Automation","author":"K\u00fcmmerle","year":"2011"},{"key":"B13","first-page":"574","article-title":"\u201cCosyPose: consistent multi-view multi-object 6D pose estimation,\u201d","volume-title":"European Conference on Computer Vision","author":"Labb\u00e9","year":"2020"},{"key":"B14","doi-asserted-by":"publisher","first-page":"239","DOI":"10.1007\/s12599-014-0334-4","article-title":"Industry 4.0","volume":"6","author":"Lasi","year":"2014","journal-title":"Bus. Inform. Syst. Eng."},{"key":"B15","doi-asserted-by":"crossref","first-page":"1785","DOI":"10.1109\/IROS.2015.7353609","article-title":"\u201cAutomatic error recovery in robot assembly operations using reverse execution,\u201d","volume-title":"2015 IEEE\/RSJ International Conference on Intelligent Robots and Systems (IROS)","author":"Laursen","year":"2015"},{"key":"B16","doi-asserted-by":"publisher","first-page":"155","DOI":"10.1007\/s11263-008-0152-6","article-title":"EPNP: an accurate O(n) solution to the PNP problem","volume":"81","author":"Lepetit","year":"2009","journal-title":"Int. J. Comput. Vis."},{"key":"B17","doi-asserted-by":"publisher","first-page":"657","DOI":"10.1007\/s11263-019-01250-9","article-title":"DeepIM: deep iterative matching for 6d pose estimation","volume":"128","author":"Li","year":"2020","journal-title":"Int. J. Comput. Vis."},{"key":"B18","doi-asserted-by":"crossref","first-page":"1150","DOI":"10.1109\/ICCV.1999.790410","article-title":"\u201cObject recognition from local scale-invariant features,\u201d","volume-title":"Proceedings of the Seventh IEEE International Conference on Computer Vision","author":"Lowe","year":"1999"},{"key":"B19","doi-asserted-by":"publisher","first-page":"205","DOI":"10.1111\/cgf.12446","article-title":"Super4PCS: fast global pointcloud registration via smart indexing","volume":"33","author":"Mellado","year":"2014","journal-title":"Comput. Graph. Forum"},{"key":"B20","doi-asserted-by":"publisher","first-page":"107193","DOI":"10.1016\/j.patcog.2019.107193","article-title":"UcoSLAM: simultaneous localization and mapping by fusion of keypoints and squared planar markers","volume":"101","author":"Munoz-Salinas","year":"2020","journal-title":"Pattern Recogn."},{"key":"B21","doi-asserted-by":"publisher","first-page":"1147","DOI":"10.1109\/TRO.2015.2463671","article-title":"ORB-SLAM: a versatile and accurate monocular SLAM system","volume":"31","author":"Mur-Artal","year":"2015","journal-title":"IEEE Trans. Robot."},{"key":"B22","first-page":"7668","article-title":"\u201cPix2pose: pixel-wise coordinate regression of objects for 6D pose estimation,\u201d","author":"Park","year":"2019","journal-title":"Proceedings of the IEEE\/CVF International Conference on Computer Vision"},{"key":"B23","doi-asserted-by":"crossref","first-page":"157","DOI":"10.1007\/3-540-44690-7_20","article-title":"\u201cServoing mechanisms for peg-in-hole assembly operations,\u201d","volume-title":"Robot Vision: International Workshop RobVis 2001 Auckland, New Zealand, February 16\u201318, 2001 Proceedings","author":"Pauli","year":"2001"},{"key":"B24","author":"Peng","year":"2019"},{"key":"B25","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1007\/s10514-017-9635-z","article-title":"Robotic assembly solution by human-in-the-loop teaching method based on real-time stiffness modulation","volume":"42","author":"Peternel","year":"2018","journal-title":"Auton. Robots"},{"key":"B26","doi-asserted-by":"publisher","first-page":"5084","DOI":"10.3390\/s19235084","article-title":"Hybrid indoor localization using IMU sensors and smartphone camera","volume":"19","author":"Poulose","year":"2019","journal-title":"Sensors"},{"key":"B27","doi-asserted-by":"publisher","DOI":"10.1109\/iccv.2017.413","article-title":"BB8: a scalable, accurate, robust to partial occlusion method for predicting the 3D poses of challenging objects without using depth","author":"Rad","year":"2017","journal-title":"arXiv preprint arXiv:1703.10896"},{"key":"B28","doi-asserted-by":"crossref","first-page":"1508","DOI":"10.1109\/ICCV.2005.104","article-title":"\u201cFusing points and lines for high performance tracking,\u201d","volume-title":"Tenth IEEE International Conference on Computer Vision (ICCV'05)","author":"Rosten","year":"2005"},{"key":"B29","doi-asserted-by":"crossref","first-page":"2564","DOI":"10.1109\/ICCV.2011.6126544","article-title":"\u201cORB: an efficient alternative to SIFT or SURF,\u201d","volume-title":"2011 International Conference on Computer Vision","author":"Rublee","year":"2011"},{"key":"B30","doi-asserted-by":"crossref","first-page":"3212","DOI":"10.1109\/ROBOT.2009.5152473","article-title":"\u201cFast point feature histograms (FPFH) for 3D registration,\u201d","volume-title":"2009 IEEE International Conference on Robotics and Automation","author":"Rusu","year":"2009"},{"key":"B31","doi-asserted-by":"publisher","first-page":"251","DOI":"10.1016\/j.cviu.2014.04.011","article-title":"SHOT: unique signatures of histograms for surface and texture description","volume":"125","author":"Salti","year":"2014","journal-title":"Comput. Vis. Image Understand."},{"key":"B32","article-title":"PCRNet: point cloud registration network using pointnet encoding","author":"Sarode","year":"2019","journal-title":"arXiv preprint arXiv:1908.07906"},{"key":"B33","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01231-1_43","article-title":"Implicit 3D orientation learning for 6D object detection from RBG images","author":"Sundermeyer","year":"2018","journal-title":"arXiv preprint arXiv:1902.01275"},{"key":"B34","first-page":"2642","article-title":"\u201cNormalized object coordinate space for category-level 6D object pose and size estimation,\u201d","author":"Wang","year":"2019","journal-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition"},{"key":"B35","first-page":"3523","article-title":"\u201cDeep closest point: learning representations for point cloud registration,\u201d","author":"Wang","year":"2019","journal-title":"Proceedings of the IEEE\/CVF International Conference on Computer Vision"},{"key":"B36","doi-asserted-by":"publisher","DOI":"10.15607\/rss.2018.xiv.019","article-title":"PoseCNN: a convolutional neural network for 6D object pose estimation in cluttered scenes","author":"Xiang","year":"2017","journal-title":"arXiv preprint arXiv:1711.00199"},{"key":"B37","doi-asserted-by":"publisher","first-page":"5072","DOI":"10.3390\/s20185072","article-title":"Multi-view-based pose estimation and its applications on intelligent manufacturing","volume":"20","author":"Yang","year":"2020","journal-title":"Sensors"},{"key":"B38","doi-asserted-by":"publisher","first-page":"2241","DOI":"10.1109\/TPAMI.2015.2513405","article-title":"Go-ICP: a globally optimal solution to 3D ICP point-set registration","volume":"38","author":"Yang","year":"2015","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"B39","first-page":"1941","article-title":"\u201cDPOD: 6D pose object detector and refiner,\u201d","author":"Zakharov","year":"2019","journal-title":"Proceedings of the IEEE\/CVF International Conference on Computer Vision"}],"container-title":["Frontiers in Neurorobotics"],"original-title":[],"link":[{"URL":"https:\/\/www.frontiersin.org\/articles\/10.3389\/fnbot.2023.1136882\/full","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2023,6,13]],"date-time":"2023-06-13T04:14:07Z","timestamp":1686629647000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.frontiersin.org\/articles\/10.3389\/fnbot.2023.1136882\/full"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,6,13]]},"references-count":39,"alternative-id":["10.3389\/fnbot.2023.1136882"],"URL":"https:\/\/doi.org\/10.3389\/fnbot.2023.1136882","relation":{},"ISSN":["1662-5218"],"issn-type":[{"value":"1662-5218","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,6,13]]},"article-number":"1136882"}}