{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,8,26]],"date-time":"2026-08-26T18:21:21Z","timestamp":1787768481524,"version":"build-2784847793"},"reference-count":90,"publisher":"American Association for the Advancement of Science (AAAS)","issue":"117","license":[{"start":{"date-parts":[[2027,8,26]],"date-time":"2027-08-26T00:00:00Z","timestamp":1819238400000},"content-version":"vor","delay-in-days":365,"URL":"https:\/\/www.science.org\/content\/page\/science-licenses-journal-article-reuse"}],"funder":[{"DOI":"10.13039\/100000181","name":"Air Force Office of Scientific Research","doi-asserted-by":"publisher","award":["FA9550-22-1-0538"],"award-info":[{"award-number":["FA9550-22-1-0538"]}],"id":[{"id":"10.13039\/100000181","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100015397","name":"Minist\u00e8re de la Culture","doi-asserted-by":"publisher","award":["RGPIN-2022-04606"],"award-info":[{"award-number":["RGPIN-2022-04606"]}],"id":[{"id":"10.13039\/100015397","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100000038","name":"Natural Sciences and Engineering Research Council of Canada","doi-asserted-by":"publisher","award":["RGPIN-2022-04606"],"award-info":[{"award-number":["RGPIN-2022-04606"]}],"id":[{"id":"10.13039\/501100000038","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100000038","name":"Natural Sciences and Engineering Research Council of Canada","doi-asserted-by":"publisher","award":["DGDND-2022-04606"],"award-info":[{"award-number":["DGDND-2022-04606"]}],"id":[{"id":"10.13039\/501100000038","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100000038","name":"Natural Sciences and Engineering Research Council of Canada","doi-asserted-by":"publisher","award":["DGDND-2022-04606"],"award-info":[{"award-number":["DGDND-2022-04606"]}],"id":[{"id":"10.13039\/501100000038","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100001804","name":"Canada Research Chairs Program","doi-asserted-by":"crossref","award":["950-231659"],"award-info":[{"award-number":["950-231659"]}],"id":[{"id":"10.13039\/501100001804","id-type":"DOI","asserted-by":"crossref"}]},{"name":"Canada Research Chairs Program Number","award":["950-231659"],"award-info":[{"award-number":["950-231659"]}]},{"name":"Natural Sciences and Engineering Research Council of Canada Number","award":["RGPIN-2022-04606"],"award-info":[{"award-number":["RGPIN-2022-04606"]}]}],"content-domain":{"domain":["www.science.org"],"crossmark-restriction":true},"short-container-title":["Sci. Robot."],"published-print":{"date-parts":[[2026,8,26]]},"abstract":"<jats:p>The design of robotic binocular camera systems has been inspired by human vision, as has their use in humanoid robots. One aspect of this inspiration has yet to play a major role, namely, that humans determine depth using a convergent binocular imaging geometry with both eyes pointing at the same location. Although some robot heads have the functionality to use vergence and version movements and thus alter binocular imaging geometry, a method to exploit it for depth computation has not been explored. Even though to an observer, their motions and appearance might resemble that of humans, in detail, the resemblance to the motions required for actual convergent depth computation is not seen. To bridge this gap, we present convergent binocular stereo (CBS), a stereo algorithm designed to provide a foundation for purposeful binocular computations under the humanoid constraint, intended for active binocular robots. CBS computes horizontal and vertical disparities using a coarse-to-fine refinement strategy with Gabor-filtered responses. We also introduce the Convergent Binocular Stereo\u2013BenchMark (CBS-BM), a convergent, natural image dataset containing 49 scenes with ground-truth horizontal disparity collected on a four-degrees-of-freedom robotic system. Our evaluation, a quantitative comparison between parallel and convergent stereo systems, shows that CBS is broadly competitive with state-of-the-art parallel methods, even outperforming them in scenes with repeated patterns and in mean horizontal disparity and depth error over all scenes. Although not intended to replace parallel stereo where humanlike behavior is unnecessary, CBS enables functionally realistic depth computation for humanoid robotic heads.<\/jats:p>","DOI":"10.1126\/scirobotics.aec7205","type":"journal-article","created":{"date-parts":[[2026,8,26]],"date-time":"2026-08-26T17:58:19Z","timestamp":1787767099000},"update-policy":"https:\/\/doi.org\/10.34133\/aaas_crossmark","source":"Crossref","is-referenced-by-count":0,"title":["Convergent binocular stereo: Depth perception for humanoid robot vision"],"prefix":"10.1126","volume":"11","author":[{"ORCID":"https:\/\/orcid.org\/0009-0002-1994-2906","authenticated-orcid":true,"given":"Mingshi","family":"Chi","sequence":"first","affiliation":[{"name":"Department of Electrical Engineering and Computer Science, York University, Toronto, Ontario M3J 1P3, Canada."},{"name":"Department of Robotics, Tohoku University, Sendai, Miyagi, Japan."}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-8621-9147","authenticated-orcid":true,"given":"John K.","family":"Tsotsos","sequence":"additional","affiliation":[{"name":"Department of Electrical Engineering and Computer Science, York University, Toronto, Ontario M3J 1P3, Canada."}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"221","reference":[{"key":"e_1_3_2_2_2","doi-asserted-by":"publisher","DOI":"10.3390\/app12167970"},{"key":"e_1_3_2_3_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.robot.2021.103834"},{"key":"e_1_3_2_4_2","doi-asserted-by":"crossref","unstructured":"M. R. M. Jenkin \u201cEvolution of robotic heads\u201d in Computer Vision: A Reference Guide K. Ikeuchi Ed. (Springer 2014) pp. 263\u2013265.","DOI":"10.1007\/978-0-387-31439-6_277"},{"key":"e_1_3_2_5_2","doi-asserted-by":"crossref","unstructured":"J. K. Tsotsos A Computational Perspective on Visual Attention (MIT Press 2011) 10.7551\/mitpress\/9780262015417.001.0001.","DOI":"10.7551\/mitpress\/9780262015417.001.0001"},{"key":"e_1_3_2_6_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41598-023-47188-4"},{"key":"e_1_3_2_7_2","doi-asserted-by":"publisher","DOI":"10.1371\/journal.pone.0319719"},{"key":"e_1_3_2_8_2","doi-asserted-by":"crossref","unstructured":"H. I. Christensen H. Bunke K. Bowyer Eds. Active Robot Vision: Camera Heads Model Based Navigation and Reactive Control vol. 6 of Series in Machine Perception and Artificial Intelligence (World Scientific Publishing Co. 1993).","DOI":"10.1142\/9789812797865"},{"key":"e_1_3_2_9_2","doi-asserted-by":"crossref","unstructured":"R. Beira M. Lopes M. Pra\u00e7a J. Santos-Victor A. Bernardino G. Metta F. Becchi R. Saltar\u00e9n \u201cDesign of the robot-cub (iCub) head\u201d in Proceedings of the IEEE International Conference on Robotics and Automation (ICRA) (IEEE 2006) pp. 94\u2013100.","DOI":"10.1109\/ROBOT.2006.1641167"},{"key":"e_1_3_2_10_2","first-page":"41","article-title":"A head-eye system\u2014Analysis and design","volume":"56","author":"Pahlavan K.","year":"1992","unstructured":"K. Pahlavan, J.-O. Eklundh, A head-eye system\u2014Analysis and design. Comput. Vis. Graph. Image Process. Image Underst. 56, 41\u201356 (1992).","journal-title":"Comput. Vis. Graph. Image Process. Image Underst."},{"key":"e_1_3_2_11_2","doi-asserted-by":"crossref","unstructured":"X. Zhang Y. Sato \u201cCooperative movements of binocular motor system\u201d in IEEE International Conference on Automation Science and Engineering (IEEE 2008) pp. 321\u2013327 10.1109\/COASE.2008.4626489.","DOI":"10.1109\/COASE.2008.4626489"},{"key":"e_1_3_2_12_2","doi-asserted-by":"publisher","DOI":"10.1142\/S0219843617500062"},{"key":"e_1_3_2_13_2","first-page":"35","article-title":"Enabling depth-driven visual attention on the iCub humanoid robot: Instructions for use and new perspectives","volume":"3","author":"Pasquale G.","year":"2016","unstructured":"G. Pasquale, T. Mar, C. Ciliberto, L. Rosasco, L. Natale, Enabling depth-driven visual attention on the iCub humanoid robot: Instructions for use and new perspectives. Front. Rob. AI 3, 35 (2016).","journal-title":"Front. Rob. AI"},{"key":"e_1_3_2_14_2","doi-asserted-by":"publisher","DOI":"10.1177\/105971230000800104"},{"key":"e_1_3_2_15_2","doi-asserted-by":"crossref","unstructured":"R. A. Brooks C. Breazeal M. Marjanovi\u0107 B. Scassellati M. M. Williamson \u201cThe COG Project: Building a humanoid robot\u201d in Computation for Metaphors Analogy and Agents C. L. Nehaniv Ed. vol. 1562 of Lecture Notes in Computer Science (Springer 1999) pp. 52\u201387.","DOI":"10.1007\/3-540-48834-0_5"},{"key":"e_1_3_2_16_2","doi-asserted-by":"publisher","DOI":"10.1007\/s10514-017-9615-3"},{"key":"e_1_3_2_17_2","doi-asserted-by":"crossref","unstructured":"D. Scharstein R. Szeliski R. Zabih \u201cA taxonomy and evaluation of dense two-frame stereo correspondence algorithms\u201d in Proceedings of the IEEE Workshop on Stereo and Multi-Baseline Vision (SMBV) (IEEE 2001) pp. 131\u2013140.","DOI":"10.1109\/SMBV.2001.988771"},{"key":"e_1_3_2_18_2","first-page":"8562323","article-title":"Review of stereo matching algorithms based on deep learning","volume":"2020","author":"Zhou K.","year":"2019","unstructured":"K. Zhou, X. Meng, B. Cheng, Review of stereo matching algorithms based on deep learning. Comput. Intell. Neurosci. 2020, 8562323 (2019).","journal-title":"Comput. Intell. Neurosci."},{"key":"e_1_3_2_19_2","first-page":"5314","article-title":"On the synergies between machine learning and binocular stereo for depth estimation from images: A survey","volume":"44","author":"Poggi M.","year":"2022","unstructured":"M. Poggi, F. Tosi, K. Batsos, P. Mordohai, S. Mattoccia, On the synergies between machine learning and binocular stereo for depth estimation from images: A survey. IEEE Trans. Pattern Anal. Mach. Intell. 44, 5314\u20135334 (2022).","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"e_1_3_2_20_2","unstructured":"R. Hartley A. Zisserman Multiple View Geometry in Computer Vision (Cambridge Univ. Press 2000)."},{"key":"e_1_3_2_21_2","unstructured":"O. Faugeras Three-Dimensional Computer Vision: A Geometric Viewpoint (MIT Press 1993)."},{"key":"e_1_3_2_22_2","doi-asserted-by":"publisher","DOI":"10.1167\/9.13.11"},{"key":"e_1_3_2_23_2","doi-asserted-by":"publisher","DOI":"10.1126\/sciadv.1400254"},{"key":"e_1_3_2_24_2","unstructured":"H. Von Helmholtz Handbuch der physiologischen Optik vol. 9 (L. Voss 1867)."},{"key":"e_1_3_2_25_2","unstructured":"E. Hering Die lehre vom binocularen sehen (Engelmann 1868)."},{"key":"e_1_3_2_26_2","doi-asserted-by":"publisher","DOI":"10.1016\/S0020-7373(75)80030-7"},{"key":"e_1_3_2_27_2","doi-asserted-by":"publisher","DOI":"10.1098\/rspb.1979.0029"},{"key":"e_1_3_2_28_2","doi-asserted-by":"publisher","DOI":"10.1007\/BF00336114"},{"key":"e_1_3_2_29_2","doi-asserted-by":"publisher","DOI":"10.1016\/1049-9660(91)90002-7"},{"key":"e_1_3_2_30_2","doi-asserted-by":"crossref","unstructured":"P. Heise S. Klose B. Jensen A. Knoll \u201cPM-Huber: PatchMatch with Huber regularization for stereo matching\u201d in Proceedings of the IEEE International Conference on Computer Vision (ICCV) (IEEE 2013) pp. 2360\u20132367.","DOI":"10.1109\/ICCV.2013.293"},{"key":"e_1_3_2_31_2","unstructured":"H. P. Moravec \u201cRobot spatial perception by stereoscopic vision and 3D evidence grids\u201d (Tech. Rep. CMU-RI-TR-96-34 Robotics Institute Carnegie Mellon University 1996)."},{"key":"e_1_3_2_32_2","doi-asserted-by":"crossref","unstructured":"D. Gallup J.-M. Frahm P. Mordohai Q. Yang M. Pollefeys \u201cReal-time plane-sweeping stereo with multiple sweeping directions\u201d in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (IEEE 2007) pp. 1\u20138 10.1109\/CVPR.2007.383245.","DOI":"10.1109\/CVPR.2007.383245"},{"key":"e_1_3_2_33_2","doi-asserted-by":"crossref","unstructured":"C. H\u00e4ne T. Sattler M. Pollefeys \u201cObstacle detection for self-driving cars using only monocular cameras and wheel odometry\u201d in Proceedings of the IEEE\/RSJ International Conference on Intelligent Robots and Systems (IROS) (IEEE 2015) pp. 5101\u20135108 10.1109\/IROS.2015.7354095.","DOI":"10.1109\/IROS.2015.7354095"},{"key":"e_1_3_2_34_2","doi-asserted-by":"crossref","unstructured":"T. Sch\u00f6ps T. Sattler C. H\u00e4ne M. Pollefeys \u201c3D modeling on the go: Interactive 3D reconstruction of large-scale scenes on mobile devices\u201d in Proceedings of the International Conference on 3D Vision (IEEE 2015) pp. 291\u2013299 10.1109\/3DV.2015.40.","DOI":"10.1109\/3DV.2015.40"},{"key":"e_1_3_2_35_2","doi-asserted-by":"crossref","unstructured":"C. H\u00e4ne L. Heng G. H. Lee A. Sizov M. Pollefeys \u201cReal-time direct dense matching on fisheye images using plane-sweeping stereo\u201d in Proceedings of the IEEE International Conference on 3D Vision vol. 1 (IEEE 2014) pp. 57\u201364 10.1109\/3DV.2014.77.","DOI":"10.1109\/3DV.2014.77"},{"key":"e_1_3_2_36_2","doi-asserted-by":"crossref","unstructured":"J. Zbontar Y. LeCun \u201cComputing the stereo matching cost with a convolutional neural network\u201d in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (IEEE 2015) pp. 1592\u20131599.","DOI":"10.1109\/CVPR.2015.7298767"},{"key":"e_1_3_2_37_2","doi-asserted-by":"crossref","unstructured":"X. Guo K. Yang W. Yang X. Wang H. Li \u201cGroup-wise correlation stereo network\u201d in Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (IEEE 2019) pp. 3268\u20133277.","DOI":"10.1109\/CVPR.2019.00339"},{"key":"e_1_3_2_38_2","doi-asserted-by":"crossref","unstructured":"L. Lipson Z. Teed J. Deng \u201cRAFT-Stereo: Multilevel recurrent field transforms for stereo matching\u201d in Proceedings of the International Conference on 3D Vision (3DV) (IEEE 2021) pp. 218\u2013227.","DOI":"10.1109\/3DV53792.2021.00032"},{"key":"e_1_3_2_39_2","doi-asserted-by":"crossref","unstructured":"F. Tosi A. Tonioni D. De Gregorio M. Poggi \u201cNeRF-supervised deep stereo\u201d in Proceedings of the IEEE\/CVF International Conference on Computer Vision and Pattern Recognition (CVPR) (IEEE 2023) pp. 855\u2013866.","DOI":"10.1109\/CVPR52729.2023.00089"},{"key":"e_1_3_2_40_2","doi-asserted-by":"crossref","unstructured":"G. Xu X. Wang X. Ding X. Yang \u201cIterative geometry encoding volume for stereo matching\u201d in Proceedings of the IEEE\/CVF International Conference on Computer Vision and Pattern Recognition (CVPR) (IEEE 2023) pp. 21919\u201321928.","DOI":"10.1109\/CVPR52729.2023.02099"},{"key":"e_1_3_2_41_2","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2025.3569218"},{"key":"e_1_3_2_42_2","doi-asserted-by":"crossref","unstructured":"Z. Chen W. Long H. Yao Y. Zhang B. Wang Y. Qin J. Wu \u201cMoCha-Stereo: Motif channel attention network for stereo matching\u201d in Proceedings of the IEEE\/CVF International Conference on Computer Vision and Pattern Recognition (CVPR) (IEEE 2024) pp. 27768\u201327777.","DOI":"10.1109\/CVPR52733.2024.02623"},{"key":"e_1_3_2_43_2","doi-asserted-by":"crossref","unstructured":"J. Li P. Wang P. Xiong T. Cai Z. Yan L. Yang J. Liu H. Fan S. Liu \u201cPractical stereo matching via cascaded recurrent network with adaptive correlation\u201d in Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (IEEE 2022) pp. 16263\u201316272.","DOI":"10.1109\/CVPR52688.2022.01578"},{"key":"e_1_3_2_44_2","doi-asserted-by":"crossref","unstructured":"H. Zhao H. Zhou Y. Zhang J. Chen Y. Yang Y. Zhao \u201cHigh-frequency stereo matching network\u201d in Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (IEEE 2023) pp. 1327\u20131336.","DOI":"10.1109\/CVPR52729.2023.00134"},{"key":"e_1_3_2_45_2","doi-asserted-by":"crossref","unstructured":"X. Wang G. Xu H. Jia X. Yang \u201cSelective-stereo: Adaptive frequency information selection for stereo matching\u201d in Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (IEEE 2024) pp. 19701\u201319710.","DOI":"10.1109\/CVPR52733.2024.01863"},{"key":"e_1_3_2_46_2","unstructured":"D. Scharstein R. Szeliski \u201cHigh-accuracy stereo depth maps using structured light\u201d in Proceedings of the IEEE\/CVF International Conference on Computer Vision and Pattern Recognition (CVPR) (IEEE 2003) 10.1109\/CVPR.2003.1211354."},{"key":"e_1_3_2_47_2","doi-asserted-by":"crossref","unstructured":"D. Scharstein C. Pal \u201cLearning conditional random fields for stereo\u201d in Proceedings of the IEEE\/CVF International Conference on Computer Vision and Pattern Recognition (CVPR) (IEEE 2007) pp. 1\u20138 10.1109\/CVPR.2007.383191.","DOI":"10.1109\/CVPR.2007.383191"},{"key":"e_1_3_2_48_2","doi-asserted-by":"crossref","unstructured":"H. Hirschmuller D. Scharstein \u201cEvaluation of cost functions for stereo matching\u201d in Proceedings of the IEEE\/CVF International Conference on Computer Vision and Pattern Recognition (CVPR) (IEEE 2007) pp. 1\u20138.","DOI":"10.1109\/CVPR.2007.383248"},{"key":"e_1_3_2_49_2","doi-asserted-by":"crossref","unstructured":"D. Scharstein H. Hirschm\u00fcller Y. Kitajima G. Krathwohl N. Ne\u0161i\u0107 X. Wang P. Westling \u201cHigh-resolution stereo datasets with subpixel-accurate ground truth\u201d in Pattern Recognition: 36th German Conference GCPR 2014 M\u00fcnster Germany September 2\u20135 2014 Proceedings X. Jiang J. Hornegger R. Koch Eds. vol. 8753 of Lecture Notes in Computer Science (Springer 2014) pp. 31\u201342.","DOI":"10.1007\/978-3-319-11752-2_3"},{"key":"e_1_3_2_50_2","doi-asserted-by":"crossref","unstructured":"T. Sch\u00f6ps J. L. Sch\u00f6nberger S. Galliani T. Sattler K. Schindler M. Pollefeys A. Geiger \u201cA multi-view stereo benchmark with high-resolution images and multi-camera videos\u201d in Proceedings of the IEEE\/CVF International Conference on Computer Vision and Pattern Recognition (CVPR) (IEEE 2017) pp. 2538\u20132547.","DOI":"10.1109\/CVPR.2017.272"},{"key":"e_1_3_2_51_2","doi-asserted-by":"crossref","unstructured":"N. Silberman D. Hoiem P. Kohli R. Fergus \u201cIndoor segmentation and support inference from RGBD images\u201d in Computer Vision \u2013 ECCV 2012: 12th European Conference on Computer Vision Florence Italy October 7\u201313 2012. Proceedings Part V A. Fitzgibbon S. Lazebnik P. Perona Y. Sato C. Schmid Eds. vol. 7576 of Lecture Notes in Computer Science (Springer 2012) pp. 746\u2013760.","DOI":"10.1007\/978-3-642-33715-4_54"},{"key":"e_1_3_2_52_2","doi-asserted-by":"crossref","unstructured":"A. Janoch S. Karayev Y. Jia J. T. Barron M. Fritz K. Saenko T. Darrell \u201cA category-level 3D object dataset: Putting the Kinect to work\u201d in Proceedings of the IEEE International Conference on Computer Vision (ICCV) (IEEE 2011) pp. 1168\u20131174.","DOI":"10.1109\/ICCVW.2011.6130382"},{"key":"e_1_3_2_53_2","doi-asserted-by":"crossref","unstructured":"J. Xiao A. Owens A. Torralba \u201cSUN3D: A database of big spaces reconstructed using SfM and object labels\u201d in Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV) (IEEE 2013) pp. 1625\u20131632.","DOI":"10.1109\/ICCV.2013.458"},{"key":"e_1_3_2_54_2","doi-asserted-by":"publisher","DOI":"10.5194\/isprsannals-II-3-W5-427-2015"},{"key":"e_1_3_2_55_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.isprsjprs.2017.09.013"},{"key":"e_1_3_2_56_2","doi-asserted-by":"crossref","unstructured":"N. Mayer E. Ilg P. H\u00e4usser P. Fischer D. Cremers A. Dosovitskiy T. Brox \u201cA large dataset to train convolutional networks for disparity optical flow and scene flow estimation\u201d in Proceedings of the IEEE\/CVF International Conference on Computer Vision and Pattern Recognition (CVPR) (2016) pp. 4040\u20134048.","DOI":"10.1109\/CVPR.2016.438"},{"key":"e_1_3_2_57_2","doi-asserted-by":"publisher","DOI":"10.1038\/sdata.2017.34"},{"key":"e_1_3_2_58_2","doi-asserted-by":"publisher","DOI":"10.1177\/2041669516681308"},{"key":"e_1_3_2_59_2","doi-asserted-by":"publisher","DOI":"10.1038\/361253a0"},{"key":"e_1_3_2_60_2","doi-asserted-by":"publisher","DOI":"10.1016\/S0042-6989(98)00139-4"},{"key":"e_1_3_2_61_2","doi-asserted-by":"crossref","unstructured":"J. G\u00e5rding T. Lindeberg \u201cDirect estimation of local surface shape in a fixating binocular vision system\u201d in Computer Vision - ECCV '94: Third European Conference on Computer Vision Stockholm Sweden May 2\u20136 1994. Proceedings Volume 1 J.-O. Eklundh Ed. vol. 800 of Lecture Notes in Computer Science (Springer 1994) pp. 365\u2013376.","DOI":"10.1007\/3-540-57956-7_40"},{"key":"e_1_3_2_62_2","doi-asserted-by":"publisher","DOI":"10.1038\/srep44800"},{"key":"e_1_3_2_63_2","doi-asserted-by":"publisher","DOI":"10.1068\/p7387"},{"key":"e_1_3_2_64_2","doi-asserted-by":"publisher","DOI":"10.1109\/LRA.2026.3682980"},{"key":"e_1_3_2_65_2","unstructured":"Itseez Open Source Computer Vision Library GitHub (2015) https:\/\/github.com\/itseez\/opencv."},{"key":"e_1_3_2_66_2","unstructured":"G. Andreas R. Martin U. Raquel \u201cEfficient large-scale stereo matching\u201d in Computer Vision - ACCV 2010: 10th Asian Conference on Computer Vision Queenstown New Zealand November 8\u201312 2010 Revised Selected Papers Part I R. Kimmel R. Klette A. Sugimoto Eds. vol. 6492 of Lecture Notes in Computer Science (2010) pp. 25\u201338."},{"key":"e_1_3_2_67_2","unstructured":"D. Ongaro J. Ousterhout \u201cIn search of an understandable consensus algorithm\u201d in Proceedings of the 2014 USENIX Conference on USENIX Annual Technical Conference (USENIX Association 2014) pp. 305\u2013320."},{"key":"e_1_3_2_68_2","doi-asserted-by":"crossref","unstructured":"N. Mayer E. Ilg P. H\u00e4usser P. Fischer D. Cremers A. Dosovitskiy T. Brox \u201cA large dataset to train convolutional networks for disparity optical flow and scene flow estimation\u201d in Proceedings of the IEEE\/CVF International Conference on Computer Vision and Pattern Recognition (CVPR) (IEEE 2016) pp. 4040\u20134048 10.1109\/CVPR.2016.438.","DOI":"10.1109\/CVPR.2016.438"},{"key":"e_1_3_2_69_2","doi-asserted-by":"crossref","unstructured":"M. D. Solbach J. K. Tsotsos \u201cBlocks world revisited: The effect of self-occlusion on classification by convolutional neural networks\u201d in Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV) Workshops (IEEE 2021) pp. 3498\u20133507.","DOI":"10.1109\/ICCVW54120.2021.00390"},{"key":"e_1_3_2_70_2","unstructured":"J. Goldman J. K. Tsotsos Statistical challenges with dataset construction: Why you will never have enough images. arXiv:2408.11160 [cs.CV] (2024)."},{"key":"e_1_3_2_71_2","unstructured":"R. Yang M. Pollefeys \u201cMulti-resolution real-time stereo on commodity graphics hardware\u201d in Proceedings of the IEEE\/CVF International Conference on Computer Vision and Pattern Recognition (CVPR) vol. 1 (IEEE 2003) 10.1109\/CVPR.2003.1211356."},{"key":"e_1_3_2_72_2","doi-asserted-by":"crossref","unstructured":"D. Gallup J.-M. Frahm P. Mordohai M. Pollefeys \u201cVariable baseline\/resolution stereo\u201d in Proceedings of the IEEE\/CVF International Conference on Computer Vision and Pattern Recognition (CVPR) (IEEE 2008) pp. 1\u20138 10.1109\/CVPR.2008.4587671.","DOI":"10.1109\/CVPR.2008.4587671"},{"key":"e_1_3_2_73_2","doi-asserted-by":"crossref","unstructured":"A. Bernardino J. Santos-Victor \u201cVergence control for robotic heads using log-polar images\u201d in Proceedings of IEEE\/RSJ International Conference on Intelligent Robots and Systems (IROS) vol. 3 (1996) pp. 1264\u20131271 10.1109\/IROS.1996.568980.","DOI":"10.1109\/IROS.1996.568980"},{"key":"e_1_3_2_74_2","doi-asserted-by":"publisher","DOI":"10.1038\/293133a0"},{"key":"e_1_3_2_75_2","doi-asserted-by":"publisher","DOI":"10.1109\/34.601246"},{"key":"e_1_3_2_76_2","doi-asserted-by":"crossref","unstructured":"D. G. Lowe \u201cObject recognition from local scale-invariant features\u201d in Proceedings of the IEEE International Conference on Computer Vision (ICCV) vol. 2 (IEEE 1999) pp. 1150\u20131157.","DOI":"10.1109\/ICCV.1999.790410"},{"key":"e_1_3_2_77_2","doi-asserted-by":"publisher","DOI":"10.1023\/A:1007984508483"},{"key":"e_1_3_2_78_2","first-page":"429","article-title":"Theory of communication","volume":"93","author":"Gabor D.","year":"1946","unstructured":"D. Gabor, Theory of communication. J. Inst. Electr. Eng. Part 93, 429\u2013457 (1946).","journal-title":"J. Inst. Electr. Eng. Part"},{"key":"e_1_3_2_79_2","doi-asserted-by":"crossref","unstructured":"F. Delattre D. Dirnfeld P. Nguyen S. Scarano M. J. Jones P. Miraldo E. Learned-Miller \u201cRobust frame-to-frame camera rotation estimation in crowded scenes\u201d in Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV) (IEEE 2023) pp. 9718\u20139728.","DOI":"10.1109\/ICCV51070.2023.00894"},{"key":"e_1_3_2_80_2","doi-asserted-by":"crossref","unstructured":"W. Xian Z. Li N. Snavely M. Fisher J. Eisenman E. Shechtman \u201cUprightNet: Geometry-aware camera orientation estimation from single images\u201d in Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV) (IEEE 2019) pp. 9973\u20139982.","DOI":"10.1109\/ICCV.2019.01007"},{"key":"e_1_3_2_81_2","doi-asserted-by":"crossref","unstructured":"A. Tsaregorodtsev J. M\u00fcller J. Strohbeck M. Herrmann M. Buchholz V. Belagiannis Extrinsic camera calibration with semantic segmentation. arXiv:2208.03949 [cs.CV] (2022).","DOI":"10.1109\/ITSC55140.2022.9922338"},{"key":"e_1_3_2_82_2","doi-asserted-by":"publisher","DOI":"10.1109\/LRA.2022.3192629"},{"key":"e_1_3_2_83_2","doi-asserted-by":"publisher","DOI":"10.1109\/70.34770"},{"key":"e_1_3_2_84_2","doi-asserted-by":"crossref","unstructured":"S. Fanello U. Pattacini I. Gori V. Tikhanoff M. Randazzo A. Roncone F. Odone G. Metta \u201c3D stereo estimation and fully automated learning of eye-hand coordination in humanoid robots\u201d in Proceedings of the IEEE-RAS International Conference on Humanoid Robots (Humanoids) (IEEE 2014) pp. 1028\u20131035.","DOI":"10.1109\/HUMANOIDS.2014.7041491"},{"key":"e_1_3_2_85_2","doi-asserted-by":"crossref","unstructured":"M. Marjanovic B. Scassellati M. Williamson \u201cSelf-taught visually-guided pointing for a humanoid robot\u201d in From Animals to Animats 4: Proceedings of the Fourth International Conference on Simulation of Adaptive Behavior J.-A. Meyer J. Pollack M. J. Mataric P. Maes S. W. Wilson Eds. (MIT Press 1996) pp. 35\u201344.","DOI":"10.7551\/mitpress\/3118.003.0007"},{"key":"e_1_3_2_86_2","doi-asserted-by":"crossref","unstructured":"L. Manfredi E. S. Maini P. Dario C. Laschi B. Girard N. Tabareau A. Berthoz \u201cImplementation of a neurophysiological model of saccadic eye movements on an anthropomorphic robotic head\u201d in Proceedings of the IEEE-RAS International Conference on Humanoid Robots (Humanoids) (IEEE 2006) pp. 438\u2013443.","DOI":"10.1109\/ICHR.2006.321309"},{"key":"e_1_3_2_87_2","doi-asserted-by":"publisher","DOI":"10.1109\/LRA.2018.2825473"},{"key":"e_1_3_2_88_2","doi-asserted-by":"publisher","DOI":"10.1109\/LRA.2024.3426293"},{"key":"e_1_3_2_89_2","doi-asserted-by":"crossref","unstructured":"W. Schenck R. M\u00f6ller \u201cStaged learning of saccadic eye movements with a robot camera head\u201d in Connectionist Models of Cognition and Perception II H. Bowman C. Labiouse Eds. vol. 15 of Progress in Neural Processing (World Scientific Publishing Co. 2004) pp. 82\u201391.","DOI":"10.1142\/9789812702784_0008"},{"key":"e_1_3_2_90_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.patcog.2014.01.005"},{"key":"e_1_3_2_91_2","doi-asserted-by":"publisher","DOI":"10.1002\/j.1538-7305.1960.tb03954.x"}],"container-title":["Science Robotics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.science.org\/doi\/pdf\/10.1126\/scirobotics.aec7205","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/www.science.org\/doi\/pdf\/10.1126\/scirobotics.aec7205","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,8,26]],"date-time":"2026-08-26T17:58:26Z","timestamp":1787767106000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.science.org\/doi\/10.1126\/scirobotics.aec7205"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,8,26]]},"references-count":90,"journal-issue":{"issue":"117","published-print":{"date-parts":[[2026,8,26]]}},"alternative-id":["10.1126\/scirobotics.aec7205"],"URL":"https:\/\/doi.org\/10.1126\/scirobotics.aec7205","relation":{},"ISSN":["2470-9476"],"issn-type":[{"value":"2470-9476","type":"electronic"}],"subject":[],"published":{"date-parts":[[2026,8,26]]},"assertion":[{"value":"2025-10-28","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2026-07-29","order":2,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2026-08-26","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}],"article-number":"eaec7205"}}