{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,23]],"date-time":"2026-07-23T09:00:52Z","timestamp":1784797252006,"version":"3.55.0"},"reference-count":64,"publisher":"Springer Science and Business Media LLC","issue":"4","license":[{"start":{"date-parts":[[2026,6,17]],"date-time":"2026-06-17T00:00:00Z","timestamp":1781654400000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/www.springernature.com\/gp\/researchers\/text-and-data-mining"},{"start":{"date-parts":[[2026,6,17]],"date-time":"2026-06-17T00:00:00Z","timestamp":1781654400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.springernature.com\/gp\/researchers\/text-and-data-mining"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Mach. Intell. Res."],"published-print":{"date-parts":[[2026,8]]},"DOI":"10.1007\/s11633-026-1654-9","type":"journal-article","created":{"date-parts":[[2026,6,17]],"date-time":"2026-06-17T06:47:46Z","timestamp":1781678866000},"page":"873-886","update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":0,"title":["Towards Understanding the Cognitive Habits of Large Reasoning Models"],"prefix":"10.1007","volume":"23","author":[{"ORCID":"https:\/\/orcid.org\/0009-0002-4576-7822","authenticated-orcid":false,"given":"Jianshuo","family":"Dong","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yujia","family":"Fu","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0008-5731-7814","authenticated-orcid":false,"given":"Chuanrui","family":"Hu","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Chao","family":"Zhang","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-2678-8070","authenticated-orcid":false,"given":"Han","family":"Qiu","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2026,6,17]]},"reference":[{"key":"1654_CR1","volume-title":"OpenAI o1 system card","author":"OpenAI","year":"2024","unstructured":"OpenAI. OpenAI o1 system card, [Online], Available: https:\/\/arxiv.org\/abs\/2412.16720, 2024."},{"key":"1654_CR2","volume-title":"DeepSeek-R1: Incentivizing reasoning capability in LLMs via reinforcement learning","author":"DeepSeek-AI","year":"2026","unstructured":"DeepSeek-AI. DeepSeek-R1: Incentivizing reasoning capability in LLMs via reinforcement learning, [Online], Available: https:\/\/arxiv.org\/abs\/2501.12948, 2026."},{"key":"1654_CR3","volume-title":"Gemini 2.5: Our Most Intelligent AI Model","author":"G DeepMind","year":"2025","unstructured":"G. DeepMind. Gemini 2.5: Our Most Intelligent AI Model, [Online], Available: https:\/\/blog.google\/innovation-and-ai\/models-and-research\/google-deepmind\/gemini-model-thinking-updates-march-2025\/, 2025."},{"key":"1654_CR4","volume-title":"s1: Simple test-time scaling","author":"N Muennighoff","year":"2025","unstructured":"N. Muennighoff, Z. Yang, W. Shi, X. L. Li, F.-F. Li, H. Hajishirzi, L. Zettlemoyer, P. Liang, E. Cand\u00e8s, T. Hashimoto. s1: Simple test-time scaling, [Online], Available: https:\/\/arxiv.org\/abs\/2501.19393, 2025."},{"key":"1654_CR5","volume-title":"Do NOT think that much for 2 + 3 = ? On the overthinking of o1-like LLMs","author":"X Chen","year":"2025","unstructured":"X. Chen, J. Xu, T. Liang, Z. He, J. Pang, D. Yu, L. Song, Q. Liu, M. Zhou, Z. Zhang, R. Wang, Z. Tu, H. Mi, D. Yu. Do NOT think that much for 2 + 3 = ? On the overthinking of o1-like LLMs, [Online], Available: https:\/\/arxiv.org\/abs\/2412.21187, 2025."},{"key":"1654_CR6","volume-title":"How should we enhance the safety of large reasoning models: An empirical study","author":"Z Zhang","year":"2025","unstructured":"Z. Zhang, X. Q. Loye, V. S. J. Huang, J. Yang, Q. Zhu, S. Cui, F. Mi, L. Shang, Y. Wang, H. Wang, M. Huang. How should we enhance the safety of large reasoning models: An empirical study, [Online], Available: https:\/\/arxiv.org\/abs\/2505.15404, 2025."},{"key":"1654_CR7","volume-title":"Monitoring reasoning models for misbehavior and the risks of promoting obfuscation","author":"B Baker","year":"2025","unstructured":"B. Baker, J. Huizinga, L. Gao, Z. Dou, M. Y. Guan, A. Madry, W. Zaremba, J. Pachocki, D. Farhi. Monitoring reasoning models for misbehavior and the risks of promoting obfuscation, [Online], Available: https:\/\/arxiv.org\/abs\/2503.11926, 2025."},{"key":"1654_CR8","volume-title":"Reasoning models don\u2019t always say what they think","author":"Y Chen","year":"2025","unstructured":"Y. Chen, J. Benton, A. Radhakrishnan, J. Uesato, C. Denison, J. Schulman, A. Somani, P. Hase, M. Wagner, F. Roger, V. Mikulik, S. R. Bowman, J. Leike, J. Kaplan, E. Perez. Reasoning models don\u2019t always say what they think, [Online], Available: https:\/\/arxiv.org\/abs\/2505.05410, 2025."},{"key":"1654_CR9","volume-title":"B. 16 habits of mind for success in learning and in life","author":"A Costa","year":"2006","unstructured":"A. Costa, B. 16 habits of mind for success in learning and in life, [Online], Available: https:\/\/habitsofmindinstitute.org\/, 2006."},{"key":"1654_CR10","volume-title":"Proceedings of the 1st Neural Information Processing Systems Track on Datasets and Benchmarks","author":"D Hendrycks","year":"2021","unstructured":"D. Hendrycks, C. Burns, S. Kadavath, A. Arora, S. Basart, E. Tang, D. Song, J. Steinhardt. Measuring mathematical problem solving with the MATH dataset. In Proceedings of the 1st Neural Information Processing Systems Track on Datasets and Benchmarks, 2021."},{"key":"1654_CR11","volume-title":"Proceedings of the 37th International Conference on Neural Information Processing Systems","author":"L Zheng","year":"2023","unstructured":"L. Zheng, W. L. Chiang, Y. Sheng, S. Zhuang, Z. Wu, Y. Zhuang, Z. Lin, Z. Li, D. Li, E. P. Xing, H. Zhang, J. E. Gonzalez, I. Stoica. Judging LLM-as-a-judge with MT-bench and Chatbot arena. In Proceedings of the 37th International Conference on Neural Information Processing Systems, New Orleans, USA, Article number 2020, 2023."},{"key":"1654_CR12","volume-title":"Siren\u2019s song in the AI ocean: A survey on hallucination in large language models","author":"Y Zhang","year":"2025","unstructured":"Y. Zhang, Y. Li, L. Cui, D. Cai, L. Liu, T. Fu, X. Huang, E. Zhao, Y. Zhang, C. Xu, Y. Chen, L. Wang, A. T. Luu, W. Bi, F. Shi, S. Shi. Siren\u2019s song in the AI ocean: A survey on hallucination in large language models, [Online], Available: https:\/\/arxiv.org\/abs\/2309.01219, 2025."},{"key":"1654_CR13","volume-title":"Proceedings of the 41st International Conference on Machine Learning","author":"M Mazeika","year":"2024","unstructured":"M. Mazeika, L. Phan, X. Yin, A. Zou, Z. Wang, N. Mu, E. Sakhaee, N. Li, S. Basart, B. Li, D. A. Forsyth, D. Hendrycks. HarmBench: A standardized evaluation framework for automated red teaming and robust refusal. In Proceedings of the 41st International Conference on Machine Learning, Vienna, Austria, Article number 1431, 2024."},{"key":"1654_CR14","volume-title":"Proceedings of the 36th International Conference on Neural Information Processing Systems","author":"J Wei","year":"2022","unstructured":"J. Wei, X. Wang, D. Schuurmans, M. Bosma, B. Ichter, F. Xia, E. H. Chi, Q. V. Le, D. Zhou. Chain-of-thought prompting elicits reasoning in large language models. In Proceedings of the 36th International Conference on Neural Information Processing Systems, New Orleans, USA, Article number 1800, 2022."},{"key":"1654_CR15","volume-title":"Improving language understanding by generative pre-training","author":"A Radford","year":"2018","unstructured":"A. Radford, K. Narasimhan. Improving language understanding by generative pre-training, [Online], Available: https:\/\/www.anthropic.com\/news\/claude-3-7-sonnet, 2018."},{"key":"1654_CR16","volume-title":"Proceedings of the 36th International Conference on Neural Information Processing Systems","author":"T Kojima","year":"2022","unstructured":"T. Kojima, S. S. Gu, M. Reid, Y. Matsuo, Y. Iwasawa. Large language models are zero-shot reasoners. In Proceedings of the 36th International Conference on Neural Information Processing Systems, New Orleans, USA, Article number 1613, 2022."},{"key":"1654_CR17","volume-title":"Proceedings of the 11th International Conference on Learning Representations","author":"F Shi","year":"2023","unstructured":"F. Shi, M. Suzgun, M. Freitag, X. Wang, S. Srivats, S. Vosoughi, H. W. Chung, Y. Tay, S. Ruder, D. Zhou, D. Das, J. Wei. Language models are multilingual chain-of-thought reasoners. In Proceedings of the 11th International Conference on Learning Representations, Kigali, Rwanda, 2023."},{"key":"1654_CR18","volume-title":"Proceedings of the 36th International Conference on Neural Information Processing Systems","author":"E Zelikman","year":"2022","unstructured":"E. Zelikman, Y. Wu, J. Mu, N. D. Goodman. STaR: Self-taught reasoner bootstrapping reasoning with reasoning. In Proceedings of the 36th International Conference on Neural Information Processing Systems, New Orleans, USA, Article number 1126, 2022."},{"key":"1654_CR19","volume-title":"Proceedings of the 12th International Conference on Learning Representations","author":"L Yu","year":"2024","unstructured":"L. Yu, W. Jiang, H. Shi, J. Yu, Z. Liu, Y. Zhang, J. T. Kwok, Z. Li, A. Weller, W. Liu. MetaMath: Bootstrap your own mathematical questions for large language models. In Proceedings of the 12th International Conference on Learning Representations, Vienna, Austria, 2024."},{"key":"1654_CR20","volume-title":"Scaling LLM testtime compute optimally can be more effective than scaling model parameters","author":"C Snell","year":"2024","unstructured":"C. Snell, J. Lee, K. Xu, A. Kumar. Scaling LLM testtime compute optimally can be more effective than scaling model parameters, [Online], Available: https:\/\/arxiv.org\/abs\/2408.03314, 2024."},{"key":"1654_CR21","volume-title":"rStar-Math: Small LLMs can master math reasoning with self-evolved deep thinking","author":"X Guan","year":"2025","unstructured":"X. Guan, L. L. Zhang, Y. Liu, N. Shang, Y. Sun, Y. Zhu, F. Yang, M. Yang. rStar-Math: Small LLMs can master math reasoning with self-evolved deep thinking, [Online], Available: https:\/\/arxiv.org\/abs\/2501.04519, 2025."},{"key":"1654_CR22","volume-title":"Claude 3.7 sonnet and claude code","author":"Anthropic","year":"2025","unstructured":"Anthropic. Claude 3.7 sonnet and claude code, [Online], Available: https:\/\/openai.com\/index\/introducing-o3-and-o4-mini\/, 2025."},{"key":"1654_CR23","volume-title":"Proceedings of the 12th International Conference on Learning Representations","author":"M Sharma","year":"2024","unstructured":"M. Sharma, M. Tong, T. Korbak, D. Duvenaud, A. Askell, S. R. Bowman, E. Durmus, Z. Hatfield-Dodds, S. R. Johnston, S. Kravec, T. Maxwell, S. McCandlish, K. Ndousse, O. Rausch, N. Schiefer, D. Yan, M. Zhang, E. Perez. Towards understanding sycophancy in language models. In Proceedings of the 12th International Conference on Learning Representations, Vienna, Austria, 2024."},{"key":"1654_CR24","doi-asserted-by":"publisher","first-page":"13484","DOI":"10.18653\/v1\/2023.acl-long.754","volume-title":"Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics","author":"Y Wang","year":"2023","unstructured":"Y. Wang, Y. Kordi, S. Mishra, A. Liu, N. A. Smith, D. Khashabi, H. Hajishirzi. Self-instruct: Aligning language models with self-generated instructions. In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics, Toronto, Canada, pp. 13484\u201313508, 2023. DOI: https:\/\/doi.org\/10.18653\/v1\/2023.acl-long.754."},{"key":"1654_CR25","volume-title":"Proceedings of the 38th International Conference on Neural Information Processing Systems","author":"X Wang","year":"2024","unstructured":"X. Wang, D. Zhou. Chain-of-thought reasoning without prompting. In Proceedings of the 38th International Conference on Neural Information Processing Systems, Vancouver, Canada, Article number 2123, 2024."},{"key":"1654_CR26","volume-title":"Language models can explain neurons in language models","author":"S Bills","year":"2023","unstructured":"S. Bills, N. Cammarata, D. Mossing, H. Tillman, L. Gao, G. Goh, I. Sutskever, J. Leike, J. Wu, W. Saunders. Language models can explain neurons in language models, [Online], Available: https:\/\/openai.com\/index\/language-models-can-explain-neurons-in-language-models\/, 2023."},{"key":"1654_CR27","volume-title":"Qwen3 technical report","author":"A Yang","year":"2025","unstructured":"A. Yang, A. Li, B. Yang, B. Zhang, B. Hui, B. Zheng, B. Yu, C. Gao, C. Huang, C. Lv, C. Zheng, D. Liu, F. Zhou, F. Huang, F. Hu, H. Ge, H. Wei, H. Lin, J. Tang, J. Yang, J. Tu, J. Zhang, J. Yang, J. Yang, J. Zhou, J. Zhou, J. Lin, K. Dang, K. Bao, K. Yang, L. Yu, L. Deng, M. Li, M. Xue, M. Li, P. Zhang, P. Wang, Q. Zhu, R. Men, R. Gao, S. Liu, S. Luo, T. Li, T. Tang, W. Yin, X. Ren, X. Wang, X. Zhang, X. Ren, Y. Fan, Y. Su, Y. Zhang, Y. Zhang, Y. Wan, Y. Liu, Z. Wang, Z. Cui, Z. Zhang, Z. Zhou, Z. Qiu. Qwen3 technical report, [Online], Available: https:\/\/arxiv.org\/abs\/2505.09388, 2025."},{"key":"1654_CR28","volume-title":"QwQ-32B: Embracing the power of reinforcement learning","author":"Qwen","year":"2025","unstructured":"Qwen. QwQ-32B: Embracing the power of reinforcement learning, [Online], Available: https:\/\/model-spec.openai.com\/2025-02-12.html, 2025."},{"key":"1654_CR29","volume-title":"Introducing OpenAI o3 and o4-Mini","author":"OpenAI","year":"2025","unstructured":"OpenAI. Introducing OpenAI o3 and o4-Mini, [Online], Available: https:\/\/qwenlm.github.io\/blog\/qwq-32b\/, 2025."},{"key":"1654_CR30","volume-title":"Seed1.5-Thinking: Advancing superb reasoning models with reinforcement learning","author":"ByteDance Seed","year":"2025","unstructured":"ByteDance Seed. Seed1.5-Thinking: Advancing superb reasoning models with reinforcement learning, [Online], Available: https:\/\/arxiv.org\/abs\/2504.13914, 2025."},{"key":"1654_CR31","volume-title":"DeepSeek-V3 technical report","author":"DeepSeek-AI","year":"2025","unstructured":"DeepSeek-AI. DeepSeek-V3 technical report, [Online], Available: https:\/\/arxiv.org\/abs\/2412.19437, 2025."},{"key":"1654_CR32","volume-title":"GPT-4o system card","author":"OpenAI","year":"2024","unstructured":"OpenAI. GPT-4o system card, [Online], Available: https:\/\/arxiv.org\/abs\/2410.21276, 2024."},{"key":"1654_CR33","volume-title":"Qwen2.5 technical report","author":"Qwen Team","year":"2025","unstructured":"Qwen Team. Qwen2.5 technical report, [Online], Available: https:\/\/arxiv.org\/abs\/2412.15115, 2025."},{"key":"1654_CR34","doi-asserted-by":"publisher","first-page":"611","DOI":"10.1145\/3600006.3613165","volume-title":"Proceedings of the 29th Symposium on Operating Systems Principles","author":"W Kwon","year":"2023","unstructured":"W. Kwon, Z. Li, S. Zhuang, Y. Sheng, L. Zheng, C. H. Yu, J. Gonzalez, H. Zhang, I. Stoica. Efficient memory management for large language model serving with PagedAttention. In Proceedings of the 29th Symposium on Operating Systems Principles, ACM, Koblenz, Germany, pp.611\u2013626, 2023."},{"key":"1654_CR35","volume-title":"Proceedings of the 37th International Conference on Neural Information Processing Systems","author":"H Liu","year":"2023","unstructured":"H. Liu, C. Li, Q. Wu, Y. J. Lee. Visual instruction tuning. In Proceedings of the 37th International Conference on Neural Information Processing Systems, New Orleans, USA, Article number 1516, 2023."},{"key":"1654_CR36","volume-title":"Proceedings of the 42nd International Conference on Machine Learning","author":"H Fei","year":"2025","unstructured":"H. Fei, Y. Zhou, J. Li, X. Li, Q. Xu, B. Li, S. Wu, Y. Wang, J. Zhou, J. Meng, Q. Shi, Z. Zhou, L. Shi, M. Gao, D. Zhang, Z. Ge, S. Tang, K. Pan, Y. Ye, H. Yuan, T. Zhang, W. Wu, T. Ju, Z. Meng, S. Xu, L. Jia, W. Hu, M. Luo, J. Luo, T. S. Chua, S. Yan, H. Zhang. On path to multimodal generalist: General-level and general-bench. In Proceedings of the 42nd International Conference on Machine Learning, Vancouver, Canada, Article number 635, 2025."},{"key":"1654_CR37","doi-asserted-by":"publisher","first-page":"6577","DOI":"10.18653\/v1\/2024.naacl-long.366","volume-title":"Proceedings of Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies","author":"J Geng","year":"2024","unstructured":"J. Geng, F. Cai, Y. Wang, H. Koeppl, P. Nakov, I. Gurevych. A survey of confidence estimation and calibration in large language models. In Proceedings of Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, ACL, Mexico City, Mexico, pp. 6577\u20136595, 2024. DOI: https:\/\/doi.org\/10.18653\/v1\/2024.naacl-long.366."},{"issue":"3","key":"1654_CR38","doi-asserted-by":"publisher","first-page":"274","DOI":"10.1007\/s00357-014-9161-z","volume":"31","author":"F Murtagh","year":"2014","unstructured":"F. Murtagh, P. Legendre. Ward\u2019s hierarchical agglomerative clustering method: Which algorithms implement ward\u2019s criterion? Journal of Classification, vol. 31, no. 3, pp. 274\u2013295, 2014. DOI: https:\/\/doi.org\/10.1007\/s00357-014-9161-z.","journal-title":"Journal of Classification"},{"key":"1654_CR39","volume-title":"Thoughts are all over the place: On the underthinking of o1-like LLMs","author":"Y Wang","year":"2025","unstructured":"Y. Wang, Q. Liu, J. Xu, T. Liang, X. Chen, Z. He, L. Song, D. Yu, J. Li, Z. Zhang, R. Wang, Z. Tu, H. Mi, D. Yu. Thoughts are all over the place: On the underthinking of o1-like LLMs, [Online], Available: https:\/\/arxiv.org\/abs\/2501.18585, 2025."},{"key":"1654_CR40","volume-title":"OpenAI model spec","author":"OpenAI","year":"2025","unstructured":"OpenAI. OpenAI model spec, [Online], Available: https:\/\/openai.com\/index\/language-unsupervised\/, 2025."},{"key":"1654_CR41","volume-title":"Proceedings of the 36th International Conference on Neural Information Processing Systems","author":"L Ouyang","year":"2022","unstructured":"L. Ouyang, J. Wu, X. Jiang, D. Almeida, C. L. Wainwright, P. Mishkin, C. Zhang, S. Agarwal, K. Slama, A. Ray, J. Schulman, J. Hilton, F. Kelton, L. Miller, M. Simens, A. Askell, P. Welinder, P. Christiano, J. Leike, R. Lowe. Training language models to follow instructions with human feedback. In Proceedings of the 36th International Conference on Neural Information Processing Systems, New Orleans, USA, Article number 2011, 2022."},{"key":"1654_CR42","volume-title":"Constitutional AI: Harmlessness from AI feedback","author":"Y Bai","year":"2022","unstructured":"Y. Bai, S. Kadavath, S. Kundu, A. Askell, J. Kernion, A. Jones, A. Chen, A. Goldie, A. Mirhoseini, C. McKinnon, C. Chen, C. Olsson, C. Olah, D. Hernandez, D. Drain, D. Ganguli, D. Li, E. Tran-Johnson, E. Perez, J. Kerr, J. Mueller, J. Ladish, J. Landau, K. Ndousse, K. Lukosuite, L. Lovitt, M. Sellitto, N. Elhage, N. Schiefer, N. Mercado, N. DasSarma, R. Lasenby, R. Larson, S. Ringer, S. Johnston, S. Kravec, S. El Showk, S. Fort, T. Lanham, T. Telleen-Lawton, T. Conerly, T. Henighan, T. Hume, S. R. Bowman, Z. Hatfield-Dodds, B. Mann, D. Amodei, N. Joseph, S. McCandlish, T. Brown, J. Kaplan. Constitutional AI: Harmlessness from AI feedback, [Online], Available: https:\/\/arxiv.org\/abs\/2212.08073, 2022."},{"issue":"5","key":"1654_CR43","doi-asserted-by":"publisher","first-page":"888","DOI":"10.1007\/s11633-024-1502-8","volume":"21","author":"T Sun","year":"2024","unstructured":"T. Sun, X. Zhang, Z. He, P., Li, Q. Cheng, X. Liu, H. Yan, Y. Shao, Q. Tang, S. Zhang, X. Zhao, K. Chen, Y. Zheng, Z. Zhou, R. Li, J. Zhan, Y. Zhou, L. Li, X. Yang, L. Wu, Z. Yin, X. Huang, Y.-G. Jiang, X. Qiu. MOSS: An open conversational large language model. Machine Intelligence Research, vol. 21, no. 5, pp. 888\u2013905, 2024. DOI: https:\/\/doi.org\/10.1007\/s11633-024-1502-8.","journal-title":"Machine Intelligence Research"},{"key":"1654_CR44","volume-title":"Proceedings of the 36th International Conference on Neural Information Processing Systems","author":"E Jones","year":"2022","unstructured":"E. Jones, J. Steinhardt. Capturing failures of large language models via human cognitive biases. In Proceedings of the 36th International Conference on Neural Information Processing Systems, New Orleans, USA, Article number 856, 2022."},{"key":"1654_CR45","volume-title":"CBEval: A framework for evaluating and interpreting cognitive biases in LLMs","author":"A Shaikh","year":"2024","unstructured":"A. Shaikh, R. A. Dandekar, S. Panat, R. Dandekar. CBEval: A framework for evaluating and interpreting cognitive biases in LLMs, [Online], Available: https:\/\/arxiv.org\/abs\/2412.03605, 2024."},{"key":"1654_CR46","doi-asserted-by":"publisher","first-page":"5269","DOI":"10.18653\/v1\/2023.findings-acl.324","volume-title":"Proceedings of Association for Computational Linguistics","author":"R Lin","year":"2023","unstructured":"R. Lin, H. T. Ng. Mind the biases: Quantifying cognitive biases in language model prompting. In Proceedings of Association for Computational Linguistics, ACL, Toronto, Canada, pp. 5269\u20135281, 2023. DOI: https:\/\/doi.org\/10.18653\/v1\/2023.findings-acl.324."},{"key":"1654_CR47","doi-asserted-by":"publisher","first-page":"27066","DOI":"10.18653\/v1\/2025.acl-long.1314","volume-title":"Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics","author":"Q Zhang","year":"2025","unstructured":"Q. Zhang, D. Wang, H. Qian, Y. Li, T. Zhang, M. Huang, K. Xu, H. Li, L. Yan, H. Qiu. Understanding the dark side of LLMs\u2019 intrinsic self-correction. In Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics, ACL, Vienna, Austria, pp. 27066\u201327101, 2025. DOI: https:\/\/doi.org\/10.18653\/v1\/2025.acl-long.1314."},{"key":"1654_CR48","volume-title":"Do LLMs possess a personality? Making the MBTI test an amazing evaluation for large language models","author":"K Pan","year":"2023","unstructured":"K. Pan, Y. Zeng. Do LLMs possess a personality? Making the MBTI test an amazing evaluation for large language models, [Online], Available: https:\/\/arxiv.org\/abs\/2307.16180, 2023."},{"key":"1654_CR49","doi-asserted-by":"publisher","first-page":"14322","DOI":"10.18653\/v1\/2024.acl-long.773","volume-title":"Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics","author":"Y Zeng","year":"2024","unstructured":"Y. Zeng, H. Lin, J. Zhang, D. Yang, R. Jia, W. Shi. How Johnny can persuade LLMs to jailbreak them: Rethinking persuasion to challenge AI safety by humanizing LLMs. In Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics, ACL, Bangkok, Thailand, pp. 14322\u201314350, 2024. DOI: https:\/\/doi.org\/10.18653\/v1\/2024.acl-long.773."},{"key":"1654_CR50","doi-asserted-by":"publisher","first-page":"16259","DOI":"10.18653\/v1\/2024.acl-long.858","volume-title":"Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics","author":"R Xu","year":"2024","unstructured":"R. Xu, B. Lin, S. Yang, T. Zhang, W. Shi, T. Zhang, Z. Fang, W. Xu, H. Qiu. The earth is flat because Investigating LLMs\u2019 belief towards misinformation via persuasive conversation. In Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics, ACL, Bangkok, Thailand, pp. 16259\u201316303, 2024. DOI: https:\/\/doi.org\/10.18653\/v1\/2024.acl-long.858."},{"key":"1654_CR51","volume-title":"AI awareness","author":"X Li","year":"2025","unstructured":"X. Li, H. Shi, R. Xu, W. Xu. AI awareness, [Online], Available: https:\/\/arxiv.org\/abs\/2504.20084, 2025."},{"key":"1654_CR52","doi-asserted-by":"publisher","first-page":"1622","DOI":"10.18653\/v1\/2024.emnlp-industry.119","volume-title":"Proceedings of Conference on Empirical Methods in Natural Language Processing: Industry Track","author":"R Xu","year":"2024","unstructured":"R. Xu, Y. Cai, Z. Zhou, R. Gu, H. Weng, L. Yan, T. Zhang, W. Xu, H. Qiu. Course-correction: Safety alignment using synthetic preferences. In Proceedings of Conference on Empirical Methods in Natural Language Processing: Industry Track, ACL, Miami, USA, pp. 1622\u20131649, 2024. DOI: https:\/\/doi.org\/10.18653\/v1\/2024.emnlp-industry.119."},{"issue":"2","key":"1654_CR53","doi-asserted-by":"publisher","first-page":"217","DOI":"10.1007\/s11633-023-1416-x","volume":"21","author":"B Cao","year":"2024","unstructured":"B. Cao, H. Lin, X. Han, L. Sun. The life cycle of knowledge in big language models: A survey. Machine Intelligence Research, vol. 21, no. 2 pp. 217\u2013238, 2024. DOI: https:\/\/doi.org\/10.1007\/s11633-023-1416-x.","journal-title":"Machine Intelligence Research"},{"issue":"3","key":"1654_CR54","doi-asserted-by":"publisher","first-page":"417","DOI":"10.1007\/s11633-025-1546-4","volume":"22","author":"Y Zhao","year":"2025","unstructured":"Y. Zhao, R. Zhang, W. Li, L. Li. Assessing and understanding creativity in large language models. Machine Intelligence Research, vol. 22, no. 3 pp. 417\u2013436, 2025. DOI: https:\/\/doi.org\/10.1007\/s11633-025-1546-4.","journal-title":"Machine Intelligence Research"},{"key":"1654_CR55","volume-title":"Cognitive behaviors that enable self-improving reasoners, or, four habits of highly effective stars","author":"K Gandhi","year":"2025","unstructured":"K. Gandhi, A. Chakravarthy, A. Singh, N. Lile, N. D. Goodman. Cognitive behaviors that enable self-improving reasoners, or, four habits of highly effective stars, [Online], Available: https:\/\/arxiv.org\/abs\/2503.01307, 2025."},{"key":"1654_CR56","volume-title":"Towards large reasoning models: A survey of reinforced reasoning with large language models","author":"F Xu","year":"2025","unstructured":"F. Xu, Q. Hao, Z. Zong, J. Wang, Y. Zhang, J. Wang, X. Lan, J. Gong, T. Ouyang, F. Meng, C. Shao, Y. Yan, Q. Yang, Y. Song, S. Ren, X. Hu, Y. Li, J. Feng, C. Gao, Y. Li. Towards large reasoning models: A survey of reinforced reasoning with large language models, [Online], Available: https:\/\/arxiv.org\/abs\/2501.09686, 2025."},{"issue":"3","key":"1654_CR57","doi-asserted-by":"publisher","first-page":"3335","DOI":"10.1109\/TPAMI.2025.3637037","volume":"48","author":"D Zhang","year":"2026","unstructured":"D. Zhang, Z. Z. Li, M. L. Zhang, J. Zhang, Z. Liu, Y. Yao, H. Xu, J. Zheng, X. Chen, Y. Zhang, F. Yin, J. Dong, Z. Guo, L. Song, C. L. Liu. From system 1 to system 2: A survey of reasoning large language models. IEEE Transactions on Pattern Analysis and Machine Intelligence, vol. 48, no. 3, pp. 3335\u20133354, 2026. DOI: https:\/\/doi.org\/10.1109\/TPAMI.2025.3637037.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"1654_CR58","volume-title":"Does reinforcement learning really incentivize reasoning capacity in LLMs beyond the base model?","author":"Y Yue","year":"2025","unstructured":"Y. Yue, Z. Chen, R. Lu, A. Zhao, Z. Wang, Y. Yue, S. Song, G. Huang. Does reinforcement learning really incentivize reasoning capacity in LLMs beyond the base model? [Online], Available: https:\/\/arxiv.org\/abs\/2504.13837, 2025."},{"key":"1654_CR59","volume-title":"Stop overthinking: A survey on efficient reasoning for large language models","author":"Y Sui","year":"2025","unstructured":"Y. Sui, Y. N. Chuang, G. Wang, J. Zhang, T. Zhang, J. Yuan, H. Liu, A. Wen, S. Zhong, N. Zou, H. Chen, X. Hu. Stop overthinking: A survey on efficient reasoning for large language models, [Online], Available: https:\/\/arxiv.org\/abs\/2503.16419, 2025."},{"issue":"3","key":"1654_CR60","doi-asserted-by":"publisher","first-page":"571","DOI":"10.1007\/s11633-024-1531-3","volume":"22","author":"Y Zhang","year":"2025","unstructured":"Y. Zhang, X.-s. Tang, K. Hao. CRMR: A collaborative multi-step reasoning framework for solving mathematical problems. Machine Intelligence Research, vol. 22, no. 3, pp. 571\u2013584, 2025. DOI: https:\/\/doi.org\/10.1007\/s11633-024-1531-3.","journal-title":"Machine Intelligence Research"},{"key":"1654_CR61","volume-title":"Detection and mitigation of hallucination in large reasoning models: A mechanistic perspective","author":"Z Sun","year":"2025","unstructured":"Z. Sun, Q. Wang, H. Wang, X. Zhang, J. Xu. Detection and mitigation of hallucination in large reasoning models: A mechanistic perspective, [Online], Available: https:\/\/arxiv.org\/abs\/2505.12886, 2025."},{"key":"1654_CR62","volume-title":"DeepSeekMath: Pushing the limits of mathematical reasoning in open language models","author":"Z Shao","year":"2024","unstructured":"Z. Shao, P. Wang, Q. Zhu, R. Xu, J. Song, X. Bi, H. Zhang, M. Zhang, Y. K. Li, Y. Wu, D. Guo. DeepSeekMath: Pushing the limits of mathematical reasoning in open language models, [Online], Available: https:\/\/arxiv.org\/abs\/2402.03300, 2024."},{"issue":"4","key":"1654_CR63","doi-asserted-by":"publisher","first-page":"740","DOI":"10.1007\/s11633-022-1407-3","volume":"21","author":"X Kong","year":"2024","unstructured":"X. Kong, S. Liu, L. Zhu. Toward human-centered XAI in practice: A survey. Machine Intelligence Research, vol. 21, no. 4, pp. 740\u2013770, 2024. DOI: https:\/\/doi.org\/10.1007\/s11633-022-1407-3.","journal-title":"Machine Intelligence Research"},{"key":"1654_CR64","volume-title":"Qwen2.5-VL technical report","author":"S Bai","year":"2025","unstructured":"S. Bai, K. Chen, X. Liu, J. Wang, W. Ge, S. Song, K. Dang, P. Wang, S. Wang, J. Tang, H. Zhong, Y. Zhu, M. Yang, Z. Li, J. Wan, P. Wang, W. Ding, Z. Fu, Y. Xu, J. Ye, X. Zhang, T. Xie, Z. Cheng, H. Zhang, Z. Yang, H. Xu, J. Lin. Qwen2.5-VL technical report, [Online], Available: https:\/\/arxiv.org\/abs\/2502.13923, 2025."}],"container-title":["Machine Intelligence Research"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s11633-026-1654-9.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s11633-026-1654-9","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s11633-026-1654-9.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,7,23]],"date-time":"2026-07-23T08:02:50Z","timestamp":1784793770000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s11633-026-1654-9"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,6,17]]},"references-count":64,"journal-issue":{"issue":"4","published-print":{"date-parts":[[2026,8]]}},"alternative-id":["1654"],"URL":"https:\/\/doi.org\/10.1007\/s11633-026-1654-9","relation":{},"ISSN":["2731-538X","2731-5398"],"issn-type":[{"value":"2731-538X","type":"print"},{"value":"2731-5398","type":"electronic"}],"subject":[],"published":{"date-parts":[[2026,6,17]]},"assertion":[{"value":"16 June 2025","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"11 March 2026","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"17 June 2026","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"The authors declared that they have no conflicts of interest to this work.","order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations of conflict of interest"}}]}}