{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,16]],"date-time":"2026-06-16T13:01:37Z","timestamp":1781614897437,"version":"3.54.5"},"reference-count":75,"publisher":"Springer Science and Business Media LLC","issue":"4","license":[{"start":{"date-parts":[[2025,9,6]],"date-time":"2025-09-06T00:00:00Z","timestamp":1757116800000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2025,9,6]],"date-time":"2025-09-06T00:00:00Z","timestamp":1757116800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"DOI":"10.13039\/501100004377","name":"The Hong Kong Polytechnic University","doi-asserted-by":"crossref","id":[{"id":"10.13039\/501100004377","id-type":"DOI","asserted-by":"crossref"}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Lang Resources &amp; Evaluation"],"published-print":{"date-parts":[[2025,12]]},"abstract":"<jats:title>Abstract<\/jats:title>\n                  <jats:p>Lexical semantic norms characterize each lexical concept in terms of a set of semantic features for the words of a language. They provide essential resources for behavioral, computational, and neuro-cognitive studies of language and human cognition. Recent research advocate for the need for cognitively motivated feature sets, arguing that semantic representations grounded in human cognition can facilitate cross-linguistic modeling and even enable the prediction of a word\u2019s semantic features based on its translation in another language. In this study, we present a new dataset of brain-based, Binder-style semantic norms for Chinese. Using the corresponding English dataset and the representational power of multilingual language models, we conduct systematic experiments on semantic norm prediction both within and across languages. We evaluate monolingual and English-Chinese cross-lingual norm prediction using two different methods: embedding-based regression vs. prompting with large language models. Our results show that bidirectional models from the BERT family and GPT-4 achieve a good level of accuracy, with moderate-to-high correlations with human ratings. Notably, in the cross-lingual setting, the best and the worst predicted features align with the higher and lower end of levels of human agreement when comparing norms of words between translated words. Our results support a novel computational approach for supplementing and expanding cognitive semantic norms, highlighting the potential of language models to bridge cross-linguistic semantic representations.<\/jats:p>","DOI":"10.1007\/s10579-025-09866-9","type":"journal-article","created":{"date-parts":[[2025,9,6]],"date-time":"2025-09-06T08:25:41Z","timestamp":1757147141000},"page":"3911-3937","update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":2,"title":["Multilingual prediction of semantic norms with language models: a study on English and Chinese"],"prefix":"10.1007","volume":"59","author":[{"given":"Bo","family":"Peng","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yu-yin","family":"Hsu","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Emmanuele","family":"Chersoni","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Le","family":"Qiu","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Chu-Ren","family":"Huang","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2025,9,6]]},"reference":[{"key":"9866_CR1","unstructured":"Achiam, J., Adler, S., Agarwal, S., Ahmad, L., Akkaya, I., Aleman, F.L., Almeida, D., Altenschmidt, J., Altman, S., Anadkat, S., & Avila, R. (2023). GPT-4 Technical Report. arXiv preprint arXiv:2303.08774"},{"issue":"2","key":"9866_CR2","first-page":"465","volume":"49","author":"M Apidianaki","year":"2023","unstructured":"Apidianaki, M. (2023). From word types to tokens and back: A survey of approaches to word meaning representation and interpretation. Computational Linguistics, 49(2), 465\u2013523.","journal-title":"Computational Linguistics"},{"key":"9866_CR3","doi-asserted-by":"crossref","unstructured":"Baroni, M., Dinu, G., & Kruszewski, G. (2014). Don\u2019t Count, Predict! A Systematic Comparison of Context-Counting vs. Context-Predicting Semantic Vectors. In: Proceedings of ACL.","DOI":"10.3115\/v1\/P14-1023"},{"key":"9866_CR4","unstructured":"BehnamGhader, P., Adlakha, V., Mosbach, M., Bahdanau, D., Chapados, N., & Reddy, S. (2024). LLM2Vec: Large language models are secretly powerful text encoders. In: Proceedings of COLM."},{"issue":"3\u20134","key":"9866_CR5","doi-asserted-by":"publisher","first-page":"130","DOI":"10.1080\/02643294.2016.1147426","volume":"33","author":"JR Binder","year":"2016","unstructured":"Binder, J. R., Conant, L. L., Humphries, C. J., Fernandino, L., Simons, S. B., Aguilar, M., & Desai, R. H. (2016). Toward a brain-based componential semantic representation. Cognitive Neuropsychology, 33(3\u20134), 130\u2013174.","journal-title":"Cognitive Neuropsychology"},{"key":"9866_CR6","doi-asserted-by":"publisher","first-page":"135","DOI":"10.1162\/tacl_a_00051","volume":"5","author":"P Bojanowski","year":"2017","unstructured":"Bojanowski, P., Grave, E., Joulin, A., & Mikolov, T. (2017). Enriching word vectors with subword information. Transactions of the Association for Computational Linguistics, 5, 135\u2013146.","journal-title":"Transactions of the Association for Computational Linguistics"},{"key":"9866_CR7","doi-asserted-by":"publisher","first-page":"333","DOI":"10.1016\/B978-0-12-385948-8.00020-7","volume-title":"Space, time and number in the brain","author":"L Boroditsky","year":"2011","unstructured":"Boroditsky, L. (2011). How languages construct time. In S. Dehaene & E. Brannon (Eds.), Space, time and number in the brain (pp. 333\u2013341). Elsevier."},{"issue":"1","key":"9866_CR8","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1006\/cogp.2001.0748","volume":"43","author":"L Boroditsky","year":"2001","unstructured":"Boroditsky, L. (2001). Does language shape thought? Mandarin and English speakers\u2019 conceptions of time. Cognitive Psychology, 43(1), 1\u201322.","journal-title":"Cognitive Psychology"},{"key":"9866_CR9","first-page":"1877","volume":"33","author":"T Brown","year":"2020","unstructured":"Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J. D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., & Agarwal, S. (2020). Language models are few-shot learners. Advances in Neural Information Processing Systems, 33, 1877\u20131901.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"9866_CR10","doi-asserted-by":"publisher","first-page":"1849","DOI":"10.3758\/s13428-019-01243-z","volume":"51","author":"EM Buchanan","year":"2019","unstructured":"Buchanan, E. M., Valentine, K. D., & Maxwell, N. P. (2019). English semantic feature production norms: an extended database of 4436 concepts. Behavior Research Methods, 51, 1849\u20131863.","journal-title":"Behavior Research Methods"},{"key":"9866_CR11","doi-asserted-by":"crossref","unstructured":"Bulat, L., Kiela, D., & Clark, S.C. (2016). Vision and Feature Norms: Improving Automatic Feature Norm Learning Through Cross-Modal Maps. In: Proceedings of NAACL-HLT.","DOI":"10.18653\/v1\/N16-1071"},{"issue":"2","key":"9866_CR12","doi-asserted-by":"publisher","first-page":"579","DOI":"10.1016\/j.cognition.2007.03.004","volume":"106","author":"D Casasanto","year":"2008","unstructured":"Casasanto, D., & Boroditsky, L. (2008). Time in the mind: Using space to think about time. Cognition, 106(2), 579\u2013593.","journal-title":"Cognition"},{"key":"9866_CR13","doi-asserted-by":"crossref","unstructured":"Chang, T.A., Arnett, C., Tu, Z., & Bergen, B.K. (2023). When is multilinguality a curse? Language modeling for 250 high-and low-resource languages. arXiv preprint arXiv:2311.09205","DOI":"10.18653\/v1\/2024.emnlp-main.236"},{"key":"9866_CR14","unstructured":"Chersoni, E., Xiang, R., Lu, Q., & Huang, C.-R. (2020). Automatic learning of modality exclusivity norms with crosslingual word embeddings. In: Proceedings of *SEM."},{"issue":"3","key":"9866_CR15","doi-asserted-by":"publisher","first-page":"663","DOI":"10.1162\/coli_a_00412","volume":"47","author":"E Chersoni","year":"2021","unstructured":"Chersoni, E., Santus, E., Huang, C.-R., & Lenci, A. (2021). Decoding word embeddings with brain-based semantic features. Computational Linguistics, 47(3), 663\u2013698.","journal-title":"Computational Linguistics"},{"key":"9866_CR16","unstructured":"Conneau, A., & Lample, G. (2019). Cross-lingual language model pretraining. In: Proceedings of the international conference on neural information processing systems. Curran Associates Inc., Red Hook, NY, USA."},{"key":"9866_CR17","doi-asserted-by":"crossref","unstructured":"Conneau, A., Khandelwal, K., Goyal, N., Chaudhary, V., Wenzek, G., Guzm\u00e1n, F., Grave, E., Ott, M., Zettlemoyer, L., & Stoyanov, V. (2020). Unsupervised cross-lingual representation learning at scale. In: Proceedings of ACL.","DOI":"10.18653\/v1\/2020.acl-main.747"},{"key":"9866_CR18","doi-asserted-by":"crossref","unstructured":"Conneau, A., Khandelwal, K., Goyal, N., Chaudhary, V., Wenzek, G., Guzm\u00e1n, F., Grave, E., Ott, M., Zettlemoyer, L., & Stoyanov, V. (2020). Unsupervised cross-lingual representation learning at scale. In: Proceedings of ACL.","DOI":"10.18653\/v1\/2020.acl-main.747"},{"key":"9866_CR19","doi-asserted-by":"crossref","unstructured":"De\u00a0Varda, A.G., Malik-Moraleda, S., Tuckute, G., & Fedorenko, E. (2025). Multilingual computational models reveal shared brain responses to 21 languages.","DOI":"10.1101\/2025.02.01.636044"},{"issue":"4","key":"9866_CR20","doi-asserted-by":"publisher","first-page":"1119","DOI":"10.3758\/s13428-013-0420-4","volume":"46","author":"BJ Devereux","year":"2014","unstructured":"Devereux, B. J., Tyler, L. K., Geertzen, J., & Randall, B. (2014). The centre for speech, language and the brain (CSLB) concept property norms. Behavior Research Methods, 46(4), 1119\u20131127.","journal-title":"Behavior Research Methods"},{"key":"9866_CR21","unstructured":"Devlin, J., Chang, M.-W., Lee, K., & Toutanova, K. (2019). BERT: Pre-training of deep bidirectional transformers for language understanding. In: Proceedings of NAACL."},{"key":"9866_CR22","unstructured":"Fagarasan, L., Vecchi, E.M., & Clark, S. (2015). From distributional semantics to feature norms: Grounding semantic models in human perceptual data. In: Proceedings of IWCS."},{"key":"9866_CR23","doi-asserted-by":"publisher","DOI":"10.1075\/hcp.37","volume-title":"Space and time in languages and cultures: Language, culture, and cognition","author":"L Filipovi\u0107","year":"2012","unstructured":"Filipovi\u0107, L., & Jaszczolt, K. M. (2012). Space and time in languages and cultures: Language, culture, and cognition. De Gruyter."},{"key":"9866_CR24","doi-asserted-by":"crossref","unstructured":"Flor, M.M. (2024). Three studies on predicting word concreteness with embedding vectors. In: Proceedings of the LREC-COLING workshop on cognitive aspects of the Lexicon.","DOI":"10.63317\/287jbvenf9fs"},{"key":"9866_CR25","unstructured":"Geiger, A., Lu, H., Icard, T., & Potts, C. (2021). Causal abstractions of neural networks. In: Proceedings. of NeurIPS."},{"key":"9866_CR26","unstructured":"Glasgow, K., Roos, M., Haufler, A., Chevillet, M., & Wolmetz, M. (2016). Evaluating semantic models with word-sentence relatedness. arXiv preprint arXiv:1603.07253"},{"key":"9866_CR27","unstructured":"Grattafiori, A., Dubey, A., Jauhri, A., Pandey, A., Kadian, A., Al-Dahle, A., Letman, A., Mathur, A., Schelten, A., Vaughan, A., Yang, A., Fan, A., Goyal, A., Hartshorn, A., Yang, A., Mitra, A., Sravankumar, A., Korenev, A., Hinsvark, A., & Ma, Z. (2024). The Llama 3 Herd of Models. arXiv preprint arXiv:2407.21783"},{"key":"9866_CR28","doi-asserted-by":"crossref","unstructured":"Hewitt, J., & Liang, P. (2019). Designing and interpreting probes with control tasks. arXiv preprint arXiv:1909.03368","DOI":"10.18653\/v1\/D19-1275"},{"issue":"1","key":"9866_CR29","doi-asserted-by":"publisher","first-page":"80","DOI":"10.1080\/00401706.2000.10485983","volume":"42","author":"AE Hoerl","year":"2000","unstructured":"Hoerl, A. E., & Kennard, R. W. (2000). Ridge regression: Biased estimation for nonorthogonal problems. Technometrics, 42(1), 80\u201386.","journal-title":"Technometrics"},{"key":"9866_CR30","unstructured":"Hui, B., Yang, J., Cui, Z., Yang, J., Liu, D., Zhang, L., Liu, T., Zhang, J., Yu, B., Dang, K., Yang, A., Men, R., Huang, F., Ren, X., Ren, X., Zhou, J., & Lin, J. (2024). Qwen2.5-Coder Technical Report, arXiv:2407.21783 [cs]"},{"key":"9866_CR31","volume-title":"Semantic structures","author":"R Jackendoff","year":"1992","unstructured":"Jackendoff, R. (1992). Semantic structures (Vol. 18). MIT Press."},{"key":"9866_CR32","unstructured":"Karthikeyan, K., Wang, Z., Mayhew, S., & Roth, D. (2019). Cross-lingual ability of multilingual BERT: An empirical study. arXiv preprint arXiv:1912.07840"},{"issue":"5","key":"9866_CR33","doi-asserted-by":"publisher","first-page":"719","DOI":"10.1207\/s15516709cog0000_33","volume":"29","author":"K Katja Wiemer-Hastings","year":"2005","unstructured":"Katja Wiemer-Hastings, K., & Xu, X. (2005). Content differences for abstract and concrete concepts. Cognitive Science, 29(5), 719\u2013736.","journal-title":"Cognitive Science"},{"issue":"11","key":"9866_CR34","doi-asserted-by":"publisher","first-page":"13386","DOI":"10.1111\/cogs.13386","volume":"47","author":"C Kauf","year":"2023","unstructured":"Kauf, C., Ivanova, A. A., Rambelli, G., Chersoni, E., She, J. S., Chowdhury, Z., Fedorenko, E., & Lenci, A. (2023). Event knowledge in large language models: The gap between the impossible and the unlikely. Cognitive Science, 47(11), 13386.","journal-title":"Cognitive Science"},{"key":"9866_CR35","doi-asserted-by":"crossref","unstructured":"Kivisaari, S.L., Hult\u00e9n, A., Vliet, M., Lindh-Knuutila, T., & Salmelin, R. (2023). Semantic Feature Norms: A Cross-method and Cross-language Comparison. Behavior Research Methods, 1\u201310.","DOI":"10.3758\/s13428-023-02311-1"},{"key":"9866_CR36","doi-asserted-by":"publisher","first-page":"97","DOI":"10.3758\/s13428-010-0028-x","volume":"43","author":"G Kremer","year":"2011","unstructured":"Kremer, G., & Baroni, M. (2011). A set of semantic norms for German and Italian. Behavior Research Methods, 43, 97\u2013109.","journal-title":"Behavior Research Methods"},{"key":"9866_CR37","unstructured":"Lample, G., Conneau, A., Denoyer, L., & Ranzato, M. (2017). Unsupervised machine translation using monolingual corpora only. arXiv preprint arXiv:1711.00043"},{"issue":"4","key":"9866_CR38","doi-asserted-by":"publisher","first-page":"1269","DOI":"10.1007\/s10579-021-09575-z","volume":"56","author":"A Lenci","year":"2022","unstructured":"Lenci, A., Sahlgren, M., Jeuniaux, P., Cuba Gyllensten, A., & Miliani, M. (2022). A comparative evaluation and analysis of three generations of distributional semantic models. Language Resources and Evaluation, 56(4), 1269\u20131313.","journal-title":"Language Resources and Evaluation"},{"key":"9866_CR39","doi-asserted-by":"publisher","first-page":"521","DOI":"10.1162\/tacl_a_00115","volume":"4","author":"T Linzen","year":"2016","unstructured":"Linzen, T., Dupoux, E., & Goldberg, Y. (2016). Assessing the ability of LSTMs to learn syntax-sensitive dependencies. Transactions of the Association for Computational Linguistics, 4, 521\u2013535.","journal-title":"Transactions of the Association for Computational Linguistics"},{"issue":"2","key":"9866_CR40","doi-asserted-by":"publisher","first-page":"30","DOI":"10.3390\/bdcc3020030","volume":"3","author":"D Li","year":"2019","unstructured":"Li, D., & Summers-Stay, D. (2019). Mapping distributional semantics to property morms with deep neural networks. Big Data and Cognitive Computing, 3(2), 30.","journal-title":"Big Data and Cognitive Computing"},{"key":"9866_CR41","unstructured":"Liu, A., Feng, B., Xue, B., Wang, B., Wu, B., Lu, C., Zhao, C., Deng, C., Zhang, C., Ruan, C., & Dai, D. (2024). Deepseek-V3 Technical Report. arXiv preprint arXiv:2412.19437"},{"key":"9866_CR42","doi-asserted-by":"crossref","unstructured":"Liu, N.F., Gardner, M., Belinkov, Y., Peters, M.E., & Smith, N.A. (2019). Linguistic Knowledge and Transferability of Contextual Representations. In: Proceedings of NAACL.","DOI":"10.18653\/v1\/N19-1112"},{"key":"9866_CR43","doi-asserted-by":"publisher","first-page":"1271","DOI":"10.3758\/s13428-019-01316-z","volume":"52","author":"D Lynott","year":"2020","unstructured":"Lynott, D., Connell, L., Brysbaert, M., Brand, J., & Carney, J. (2020). The Lancaster sensorimotor norms: Multidimensional measures of perceptual and action strength for 40,000 English words. Behavior Research Methods, 52, 1271\u20131291.","journal-title":"Behavior Research Methods"},{"key":"9866_CR44","doi-asserted-by":"crossref","unstructured":"Mart\u00ednez, G., Conde, J., Reviriego, P., & Brysbaert, M. (2024). AI-generated Estimates of Familiarity, Concreteness, Valence, and Arousal for over 100,000 Spanish Words. Quarterly Journal of Experimental Psychology.","DOI":"10.31234\/osf.io\/zqfsj"},{"key":"9866_CR45","doi-asserted-by":"crossref","unstructured":"Mart\u00ednez, G., Molero, J.D., Gonz\u00e1lez, S., Conde, J., Brysbaert, M., & Reviriego, P. (2024). Using large language models to estimate features of multi-word expressions: Concreteness, valence, arousal. arXiv preprint arXiv:2408.16012","DOI":"10.3758\/s13428-024-02515-z"},{"issue":"4","key":"9866_CR46","doi-asserted-by":"publisher","first-page":"547","DOI":"10.3758\/BF03192726","volume":"37","author":"K McRae","year":"2005","unstructured":"McRae, K., Cree, G. S., Seidenberg, M. S., & McNorgan, C. (2005). Semantic feature production norms for a large set of living and nonliving things. Behavior Research Methods, 37(4), 547\u2013559.","journal-title":"Behavior Research Methods"},{"key":"9866_CR47","unstructured":"Mikolov, T., Chen, K., Corrado, G., & Dean, J. (2013). Efficient Estimation of Word Representations in Vector Space. arXiv preprint arXiv:1301.3781"},{"key":"9866_CR48","doi-asserted-by":"publisher","first-page":"440","DOI":"10.3758\/s13428-012-0263-4","volume":"45","author":"M Montefinese","year":"2013","unstructured":"Montefinese, M., Ambrosini, E., Fairfield, B., & Mammarella, N. (2013). Semantic memory: A feature-based analysis and new norms for Italian. Behavior Research Methods, 45, 440\u2013461.","journal-title":"Behavior Research Methods"},{"key":"9866_CR49","volume-title":"The big book of concepts","author":"G Murphy","year":"2004","unstructured":"Murphy, G. (2004). The big book of concepts. MIT Press."},{"key":"9866_CR50","doi-asserted-by":"crossref","unstructured":"Pennington, J., Socher, R., & Manning, C. (2014). GloVe: Global vectors for word representation. In: Proceedings of EMNLP.","DOI":"10.3115\/v1\/D14-1162"},{"key":"9866_CR51","doi-asserted-by":"crossref","unstructured":"Peters, M.E., Neumann, M., Iyyer, M., Gardner, M., Clark, C., Lee, K., & Zettlemoyer, L. (2018). Deep contextualized word representations. In: Proceedings of NAACL.","DOI":"10.18653\/v1\/N18-1202"},{"key":"9866_CR52","doi-asserted-by":"crossref","unstructured":"Pires, T., Schlinger, E., & Garrette, D. (2019). How multilingual is multilingual BERT? In: Proceedings of ACL.","DOI":"10.18653\/v1\/P19-1493"},{"key":"9866_CR53","unstructured":"Qiu, L., Hsu, Y.-Y., & Chersoni, E. (2023). Collecting and predicting neurocognitive norms for Mandarin Chinese. In: Proceedings of IWCS."},{"key":"9866_CR54","unstructured":"Radford, A., Wu, J., Child, R., Luan, D., Amodei, D., & Sutskever, I. (2019). Language models are unsupervised multitask learners. OpenAI Blog."},{"key":"9866_CR55","doi-asserted-by":"crossref","unstructured":"Rivi\u00e8re, P.D., Beatty-Mart\u00ednez, A.L., & Trott, S. (2024). Evaluating contextualized representations of (Spanish) ambiguous words: A new lexical resource and empirical analysis. arXiv preprint arXiv:2406.14678","DOI":"10.18653\/v1\/2025.naacl-long.422"},{"key":"9866_CR56","unstructured":"Rogers, A. (2023). Closed AI Models Make Bad Baselines. Towards Data Science."},{"key":"9866_CR57","doi-asserted-by":"publisher","first-page":"1258","DOI":"10.3758\/s13428-018-1099-3","volume":"51","author":"GG Scott","year":"2019","unstructured":"Scott, G. G., Keitel, A., Becirspahic, M., Yao, B., & Sereno, S. C. (2019). The glasgow norms: Ratings of 5,500 words on nine scales. Behavior Research Methods, 51, 1258\u20131270.","journal-title":"Behavior Research Methods"},{"key":"9866_CR58","unstructured":"Springer, J.M., Kotha, S., Fried, D., Neubig, G., & Raghunathan, A. (2024). Repetition Improves Language Model Embeddings. arXiv preprint arXiv:2402.15449"},{"key":"9866_CR59","doi-asserted-by":"crossref","unstructured":"Tan, Z., Zhang, X., Wang, S., & Liu, Y. (2022). MSP: Multi-stage prompting for making pre-trained language models better translators. In: Proceedings of ACL.","DOI":"10.18653\/v1\/2022.acl-long.424"},{"key":"9866_CR60","unstructured":"Tenney, I., Xia, P., Chen, B., Wang, A., Poliak, A., McCoy, R.T., Kim, N., Durme, B.V., Bowman, S., Das, D., & Pavlick, E. (2019). What do you learn from context? Probing for sentence structure in contextualized word representations. In: Proceedings of ICLR."},{"key":"9866_CR61","unstructured":"Thompson, B., & Lupyan, G. (2018). Automatic estimation of lexical concreteness in 77 languages. In: Proceedings of the annual meeting of the cognitive science society, vol. 40."},{"key":"9866_CR62","unstructured":"Touvron, H., Lavril, T., Izacard, G., Martinet, X., Lachaux, M.-A., Lacroix, T., Rozi\u00e8re, B., Goyal, N., Hambro, E., Azhar, F., & Rodriguez, A. (2023). LLaMA: Open and Efficient Foundation Language Models. arXiv preprint arXiv:2302.13971"},{"key":"9866_CR63","doi-asserted-by":"crossref","unstructured":"Trott, S. (2024). Can large language models help augment english psycholinguistic datasets? Behavior research methods, 1\u201319.","DOI":"10.31234\/osf.io\/jvenz"},{"key":"9866_CR64","unstructured":"Trott, S., & Bergen, B. (2022). Contextualized Sensorimotor Norms: Multi-Dimensional Measures of Sensorimotor Strength for Ambiguous English Words in Context. arXiv preprint arXiv:2203.05648"},{"key":"9866_CR65","doi-asserted-by":"publisher","first-page":"141","DOI":"10.1613\/jair.2934","volume":"37","author":"PD Turney","year":"2010","unstructured":"Turney, P. D., & Pantel, P. (2010). From frequency to meaning: Vector space models of semantics. Journal of Artificial Intelligence Research, 37, 141\u2013188.","journal-title":"Journal of Artificial Intelligence Research"},{"issue":"6","key":"9866_CR66","doi-asserted-by":"publisher","first-page":"12844","DOI":"10.1111\/cogs.12844","volume":"44","author":"A Utsumi","year":"2020","unstructured":"Utsumi, A. (2020). Exploring what is encoded in distributional word vectors: A neurobiologically motivated analysis. Cognitive Science, 44(6), 12844.","journal-title":"Cognitive Science"},{"key":"9866_CR67","unstructured":"Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A.N., Kaiser, \u0141., & Polosukhin, I. (2017). Attention Is All You Need. In: Advances in Neural Information Processing Systems, pp. 5998\u20136008."},{"issue":"1","key":"9866_CR68","doi-asserted-by":"publisher","first-page":"183","DOI":"10.3758\/BRM.40.1.183","volume":"40","author":"DP Vinson","year":"2008","unstructured":"Vinson, D. P., & Vigliocco, G. (2008). Semantic feature production norms for a large set of objects and events. Behavior Research Methods, 40(1), 183\u2013190.","journal-title":"Behavior Research Methods"},{"key":"9866_CR69","doi-asserted-by":"publisher","first-page":"1095","DOI":"10.3758\/s13428-016-0777-2","volume":"49","author":"J Vivas","year":"2017","unstructured":"Vivas, J., Vivas, L., Comesa\u00f1a, A., Coni, A. G., & Vorano, A. (2017). Spanish Semantic Feature Production Norms for 400 Concrete Concepts. Behavior Research Methods, 49, 1095\u20131106.","journal-title":"Behavior Research Methods"},{"key":"9866_CR70","doi-asserted-by":"crossref","unstructured":"Vuli\u0107, I., Ponti, E.M., Litschko, R., Glava\u0161, G., & Korhonen, A. (2020). Probing pretrained language models for lexical semantics. In: Proceedings of EMNLP.","DOI":"10.18653\/v1\/2020.emnlp-main.586"},{"issue":"1","key":"9866_CR71","doi-asserted-by":"publisher","first-page":"106","DOI":"10.1038\/s41597-023-01995-6","volume":"10","author":"S Wang","year":"2023","unstructured":"Wang, S., Zhang, Y., Shi, W., Zhang, G., Zhang, J., Lin, N., & Zong, C. (2023). A large dataset of semantic ratings and its computational extension. Scientific Data, 10(1), 106.","journal-title":"Scientific Data"},{"key":"9866_CR72","doi-asserted-by":"publisher","DOI":"10.1093\/oso\/9780198700029.001.0001","volume-title":"Semantics: Primes and universals","author":"A Wierzbicka","year":"1996","unstructured":"Wierzbicka, A. (1996). Semantics: Primes and universals. Oxford University Press."},{"key":"9866_CR73","doi-asserted-by":"crossref","unstructured":"Wu, Z., Chen, Y., Kao, B., & Liu, Q. (2020). Perturbed masking: Parameter-free probing for analyzing and interpreting BERT. In: Proceedings of ACL.","DOI":"10.18653\/v1\/2020.acl-main.383"},{"key":"9866_CR74","unstructured":"Xu, Q., Peng, Y., Nastase, S.A., Chodorow, M., Wu, M., & Li, P. (2023). Does conceptual representation require embodiment? Insights from large language models. arXiv preprint arXiv:2305.19103"},{"key":"9866_CR75","doi-asserted-by":"crossref","unstructured":"Zhao, Z., Chen, H., Zhang, J., Zhao, X., Liu, T., Lu, W., Chen, X., Deng, H., Ju, Q., & Du, X. (2019). UER: an open-source toolkit for pre-training models. In: Proceedings of EMNLP-IJCNLP: System demonstrations.","DOI":"10.18653\/v1\/D19-3041"}],"container-title":["Language Resources and Evaluation"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s10579-025-09866-9.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s10579-025-09866-9","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s10579-025-09866-9.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,6,16]],"date-time":"2026-06-16T12:45:29Z","timestamp":1781613929000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s10579-025-09866-9"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,9,6]]},"references-count":75,"journal-issue":{"issue":"4","published-print":{"date-parts":[[2025,12]]}},"alternative-id":["9866"],"URL":"https:\/\/doi.org\/10.1007\/s10579-025-09866-9","relation":{},"ISSN":["1574-020X","1574-0218"],"issn-type":[{"value":"1574-020X","type":"print"},{"value":"1574-0218","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,9,6]]},"assertion":[{"value":"5 March 2025","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"8 July 2025","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"6 September 2025","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"The authors declare no Conflict of interest.","order":2,"name":"Ethics","group":{"name":"EthicsHeading","label":"Conflict of interest"}}]}}