{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,10]],"date-time":"2026-07-10T06:13:47Z","timestamp":1783664027772,"version":"3.55.0"},"reference-count":35,"publisher":"Cambridge University Press (CUP)","issue":"2","license":[{"start":{"date-parts":[[2022,11,15]],"date-time":"2022-11-15T00:00:00Z","timestamp":1668470400000},"content-version":"unspecified","delay-in-days":0,"URL":"http:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Robotica"],"published-print":{"date-parts":[[2023,2]]},"abstract":"<jats:title>Abstract<\/jats:title><jats:p>We propose a hierarchical cognitive navigation model (HCNM) to improve the self-learning and self-adaptive ability of mobile robots in unknown and complex environments. The HCNM model adopts the divide and conquers approach by dividing the path planning task into different levels of sub-tasks in complex environments and solves each sub-task in a smaller state subspace to decrease the state space dimensions. The HCNM model imitates animal asymptotic properties through the study of thermodynamic processes and designs a cognitive learning algorithm to achieve online optimum search strategies. We prove that the learning algorithm designed ensures that the cognitive model can converge to the optimal behavior path with probability one. Robot navigation is studied on the basis of the cognitive process. The experimental results show that the HCNM model has strong adaptability in unknown and environment, and the navigation path is clearer and the convergence time is better. Among them, the convergence time of HCNM model is 25 s, which is 86.5% lower than that of HRLM model. The HCNM model studied in this paper adopts a hierarchical structure, which reduces the learning difficulty and accelerates the learning speed in the unknown environment.<\/jats:p>","DOI":"10.1017\/s0263574722001539","type":"journal-article","created":{"date-parts":[[2022,11,15]],"date-time":"2022-11-15T12:36:08Z","timestamp":1668515768000},"page":"690-712","source":"Crossref","is-referenced-by-count":6,"title":["Autonomous robot navigation based on a hierarchical cognitive model"],"prefix":"10.1017","volume":"41","author":[{"given":"Jianxian","family":"Cai","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Fenfen","family":"Yan","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yan","family":"Shi","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-3872-7312","authenticated-orcid":false,"given":"Mengying","family":"Zhang","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Lili","family":"Guo","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"56","published-online":{"date-parts":[[2022,11,15]]},"reference":[{"key":"S0263574722001539_ref18","unstructured":"[18] Touretzky, D. S. , Daw, N. D. and Tira-Thompson, E. J. , \u201cCombining Configural and TD Learning on a Robot,\u201d In: Proceedings of 2th IEEE Conference on Development and Learning, Los Alamitos, CA, USA (2002) pp. 47\u201352."},{"key":"S0263574722001539_ref35","doi-asserted-by":"crossref","first-page":"864","DOI":"10.1016\/j.jmaa.2017.06.017","article-title":"Semi-Markov decision processes with limiting ratio average rewards","volume":"455","author":"Sagnik","year":"2017","journal-title":"J. Math. Anal. Appl."},{"key":"S0263574722001539_ref20","first-page":"178117","volume-title":"IEEE Access","volume":"8","author":"Zhang"},{"key":"S0263574722001539_ref12","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1016\/j.robot.2016.06.003","article-title":"Neural inverse reinforcement learning in autonomous navigation","volume":"84","author":"Chen","year":"2016","journal-title":"Robot. Auton. Syst."},{"key":"S0263574722001539_ref33","doi-asserted-by":"publisher","DOI":"10.1371\/journal.pone.0180234"},{"key":"S0263574722001539_ref11","first-page":"254","article-title":"Q - learning environmental cognition method based on odor reward guidance","volume":"61","author":"Ruan","year":"2021","journal-title":"J. Tsinghua Univ. (Nat. Sci. Edn.)"},{"key":"S0263574722001539_ref22","first-page":"79","article-title":"Application of biological learning theories to mobile robot avoidance and approach behaviors","volume":"1","author":"Chang","year":"1998","journal-title":"J. Compl. Syst"},{"key":"S0263574722001539_ref7","doi-asserted-by":"publisher","DOI":"10.1109\/TSMCB.2003.811769"},{"key":"S0263574722001539_ref5","doi-asserted-by":"crossref","unstructured":"[5] Pandey, A. , Sonkar, R. K. , Pandey, K. K. and Parhi, D. R. , \u201cPath Planning Navigation of Mobile Robot with Obstacles Avoidance Using Fuzzy Logic Controller,\u201d In: Proc. 8th IEEE Conference on Intelligent System and Control, Coimbatore, India (2014) pp. 39\u201341.","DOI":"10.1109\/ISCO.2014.7103914"},{"key":"S0263574722001539_ref26","first-page":"930","article-title":"Operant conditioning learning automatic and its application on robot balance control","volume":"28","author":"Gao","year":"2013","journal-title":"Control Decis"},{"key":"S0263574722001539_ref25","article-title":"Operant conditioning reflex learning control scheme based on SMC and Elman network","volume":"26","author":"Ruan","journal-title":"Control Decis"},{"key":"S0263574722001539_ref27","first-page":"29","article-title":"Behavior cognition computational model based on cerebellum and basal ganglia mechanism","volume":"25","author":"Chen","year":"2012","journal-title":"Patt. Recognit. Artif. Intell"},{"key":"S0263574722001539_ref8","doi-asserted-by":"crossref","first-page":"1279","DOI":"10.1109\/TNN.2008.2000394","article-title":"A bioinspired neural network for real-time concurrent map building and complete coverage robot navigation in unknown environments","volume":"99","author":"Luo","year":"2008","journal-title":"IEEE Trans. Neural Netw"},{"key":"S0263574722001539_ref3","unstructured":"[3] Wang, C. , Soh, Y. C. , Wang, H. and Wang, H. , \u201cA Hierarchical Genetic Algorithm for Path Planning in a Static Environment with Obstacles,\u201d In: Proc. IEEE Conference on Electrical and Computer Engineering, Honolulu, USA (2002) pp. 1652\u20131657."},{"key":"S0263574722001539_ref17","doi-asserted-by":"publisher","DOI":"10.1101\/lm.37501"},{"key":"S0263574722001539_ref19","doi-asserted-by":"publisher","DOI":"10.1007\/s12204-017-1814-8"},{"key":"S0263574722001539_ref32","first-page":"9","article-title":"Autonomous navigation of mobile robot based on cognitive development","volume":"44","author":"Cai","journal-title":"Comput. Eng."},{"key":"S0263574722001539_ref21","doi-asserted-by":"crossref","unstructured":"[21] Dominguez, S. , Zalama, E. , Garc\u00eda-Bermejo, J. and J., P. , \u201cRobot Learning in a Social Robot,\u201d In: From Animals to Animats 9, Lecture Notes in Computer Science, vol. 4095 (2006) pp. 691\u2013702.","DOI":"10.1007\/11840541_57"},{"key":"S0263574722001539_ref1","first-page":"1579","article-title":"Autonomous navigation and obstacle avoidance of a microbus","volume":"118","author":"Kanmani","year":"2013","journal-title":"J. Adv. Robot. Syst"},{"key":"S0263574722001539_ref4","first-page":"406","article-title":"Trajectory planning of an autonomous mobile robot by evolving ant colony system","volume":"32","author":"Wang","year":"2017","journal-title":"Int. J. Robot. Automat"},{"key":"S0263574722001539_ref2","first-page":"384","article-title":"TYPE-2 fuzzy logic controller design using a real-time PSO algorithm applied to \u201cIROBOT CREATE robot","volume":"30","author":"Srine","year":"2015","journal-title":"Int. J. Robot. Automat"},{"key":"S0263574722001539_ref15","first-page":"109","article-title":"Fuzzy logic and reinforcement learning based approaches for mobile robot navigation in unknown environment","volume":"9","author":"Cherroun","year":"2013","journal-title":"J. Meas. Cont"},{"key":"S0263574722001539_ref9","first-page":"413","article-title":"Autonomous land vehicle path planning based on learning classifiler system in narrow environments","volume":"40","author":"Shao","year":"2011","journal-title":"Inform. Comput."},{"key":"S0263574722001539_ref28","doi-asserted-by":"publisher","DOI":"10.1007\/s11431-013-5369-0"},{"key":"S0263574722001539_ref30","doi-asserted-by":"publisher","DOI":"10.1155\/2016\/4296356"},{"key":"S0263574722001539_ref23","doi-asserted-by":"publisher","DOI":"10.1162\/106454604322875913"},{"key":"S0263574722001539_ref13","volume-title":"Enhanced Learning and Approximate Dynamic Programming","author":"Xu","year":"2010"},{"key":"S0263574722001539_ref29","first-page":"1016","article-title":"A learning model based on operant conditioning principles","volume":"6","author":"Ruan","year":"2014","journal-title":"Control Decis"},{"key":"S0263574722001539_ref24","first-page":"1452","article-title":"Autonomous operant conditioning automata","volume":"29","author":"Ruan","year":"2012","journal-title":"Control Theory Appl"},{"key":"S0263574722001539_ref14","doi-asserted-by":"publisher","DOI":"10.1016\/j.robot.2015.04.003"},{"key":"S0263574722001539_ref16","first-page":"110","volume-title":"The Behavior of Organisms: An Experimental Analysis","author":"Skinner"},{"key":"S0263574722001539_ref10","first-page":"464","article-title":"Bioinspired neural network-based Q-learning approach for robot path planning in unknown environments","volume":"31","author":"Ni","year":"2016","journal-title":"Int. J. Robot. Autom."},{"key":"S0263574722001539_ref31","first-page":"1281","article-title":"Monocular vision slam of mobile robot based on Skinner-RANSAC","volume":"42","author":"Ruan","journal-title":"J. Beijing Univ. Technol."},{"key":"S0263574722001539_ref6","first-page":"441","article-title":"Selective topological approach to mobile robot navigation with recurrent neural networks","volume":"30","author":"Tom","year":"2015","journal-title":"Int. J. Robot. Automat"},{"key":"S0263574722001539_ref34","doi-asserted-by":"crossref","unstructured":"[34] Li, Z. , Narayan, A. and Leong, T. Y. , \u201cAn Efficient Approach to Model-Based Hierarchical Reinforcement Learning,\u201d In: Proceedings of 31th AAAI Conference on Artificial Intelligence, California, USA (2017) pp. 3583\u20133589.","DOI":"10.1609\/aaai.v31i1.11024"}],"container-title":["Robotica"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.cambridge.org\/core\/services\/aop-cambridge-core\/content\/view\/S0263574722001539","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2023,1,11]],"date-time":"2023-01-11T03:50:35Z","timestamp":1673409035000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.cambridge.org\/core\/product\/identifier\/S0263574722001539\/type\/journal_article"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,11,15]]},"references-count":35,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2023,2]]}},"alternative-id":["S0263574722001539"],"URL":"https:\/\/doi.org\/10.1017\/s0263574722001539","relation":{},"ISSN":["0263-5747","1469-8668"],"issn-type":[{"value":"0263-5747","type":"print"},{"value":"1469-8668","type":"electronic"}],"subject":[],"published":{"date-parts":[[2022,11,15]]}}}