{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,8,3]],"date-time":"2026-08-03T23:09:58Z","timestamp":1785798598842,"version":"3.56.0"},"reference-count":40,"publisher":"Oxford University Press (OUP)","issue":"4","license":[{"start":{"date-parts":[[2024,3,22]],"date-time":"2024-03-22T00:00:00Z","timestamp":1711065600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/academic.oup.com\/pages\/standard-publication-reuse-rights"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2024,7,25]]},"abstract":"<jats:title>Abstract<\/jats:title>\n               <jats:p>The multi-modal tasks have started to play a significant role in the research on artificial intelligence. A particular example of that domain is visual\u2013linguistic tasks, such as visual question answering. The progress of modern machine learning systems is determined, among other things, by the data on which these systems are trained. Most modern visual question answering data sets contain limited type questions that can be answered either by directly accessing the image itself or by using external data. At the same time, insufficient attention is paid to the issues of social interactions between people, which limits the scope of visual question answering systems. In this paper, we propose criteria by which images suitable for social interaction visual question answering can be selected for composing such questions, based on psychological research. We believe this should serve the progress of visual question answering systems.<\/jats:p>","DOI":"10.1093\/jigpal\/jzae026","type":"journal-article","created":{"date-parts":[[2024,3,24]],"date-time":"2024-03-24T03:43:37Z","timestamp":1711251817000},"page":"656-670","source":"Crossref","is-referenced-by-count":2,"title":["Sign-based image criteria for social interaction visual question answering"],"prefix":"10.1093","volume":"32","author":[{"given":"Anfisa A","family":"Chuganskaya","sequence":"first","affiliation":[{"name":"Federal Research Center \u201cComputer Science and Control\u201d of the Russian Academy of Sciences , 60th October Anniversary prospect 9, 117312, Moscow, Russia , anfisa.makh@gmail.com"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Alexey K","family":"Kovalev","sequence":"additional","affiliation":[{"name":"Artificial Intelligence Research Institute , Kutuzovsky ave. 32, b.1, 121170, Moscow, Russia ; , Kerchenskaya str. 1 A, b.1, 117303, Moscow, Russia , alexeykkov@gmail.com"},{"name":"Moscow Institute of Physics and Technology , Kutuzovsky ave. 32, b.1, 121170, Moscow, Russia ; , Kerchenskaya str. 1 A, b.1, 117303, Moscow, Russia , alexeykkov@gmail.com"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Aleksandr I","family":"Panov","sequence":"additional","affiliation":[{"name":"Artificial Intelligence Research Institute , Kutuzovsky ave. 32, b.1, 121170, Moscow, Russia ; , Kerchenskaya str. 1 A, b.1, 117303, Moscow, Russia , panov.ai@mipt.ru"},{"name":"Moscow Institute of Physics and Technology , Kutuzovsky ave. 32, b.1, 121170, Moscow, Russia ; , Kerchenskaya str. 1 A, b.1, 117303, Moscow, Russia , panov.ai@mipt.ru"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"286","published-online":{"date-parts":[[2024,3,22]]},"reference":[{"key":"2024072520071813700_ref1","volume-title":"Working Notes of CLEF 2019 - Conference and Labs of the Evaluation Forum","author":"Abacha","year":"2019"},{"key":"2024072520071813700_ref2","first-page":"2425","volume-title":"Proceedings of the IEEE International Conference on Computer Vision","author":"Agrawal","year":"2015"},{"key":"2024072520071813700_ref3","first-page":"374","article-title":"Remembering: a study in experimental and social psychology","volume":"8","author":"Bartlett","year":"1932","journal-title":"Philosophy"},{"key":"2024072520071813700_ref4","volume-title":"On Dexterity and its Development [in Russian]","author":"Bernstein","year":"1991"},{"key":"2024072520071813700_ref5","doi-asserted-by":"crossref","first-page":"97","DOI":"10.1016\/j.bica.2015.09.003","article-title":"Towards integrated neural\u2013symbolic systems for human-level ai: two research programs helping to bridge the gaps","volume":"14","author":"Besold","year":"2015","journal-title":"Biologically Inspired Cognitive Architectures"},{"key":"2024072520071813700_ref6","first-page":"1080","volume-title":"IEEE Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Das","year":"2017"},{"key":"2024072520071813700_ref7","volume-title":"The Perception of the Visual World","author":"Gibson","year":"1950"},{"key":"2024072520071813700_ref8","first-page":"6325","volume-title":"IEEE Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Goyal","year":"2017"},{"key":"2024072520071813700_ref9","first-page":"939","volume-title":"IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Gurari","year":"2019"},{"key":"2024072520071813700_ref10","first-page":"3608","volume-title":"IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"Gurari","year":"2018"},{"key":"2024072520071813700_ref11","doi-asserted-by":"crossref","first-page":"335","DOI":"10.1016\/0167-2789(90)90087-6","article-title":"The symbol grounding problem","volume":"42","author":"Harnad","year":"1990","journal-title":"Physica D: Nonlinear Phenomena"},{"key":"2024072520071813700_ref12","volume-title":"Working Notes of CLEF 2018 - Conference and Labs of the Evaluation Forum","author":"Hasan","year":"2018"},{"key":"2024072520071813700_ref13","first-page":"770","volume-title":"IEEE Conference on Computer Vision and Pattern Recognition (CVPR)","author":"He","year":"2016"},{"key":"2024072520071813700_ref14","first-page":"6693","volume-title":"IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Hudson","year":"2019"},{"key":"2024072520071813700_ref15","first-page":"1988","volume-title":"IEEE Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Johnson","year":"2017"},{"key":"2024072520071813700_ref16","doi-asserted-by":"crossref","first-page":"60","DOI":"10.18653\/v1\/2020.bionlp-1.6","volume-title":"Proceedings of the 19th SIGBioMed Workshop on Biomedical Language Processing","author":"Kovaleva","year":"2020"},{"key":"2024072520071813700_ref17","doi-asserted-by":"crossref","first-page":"32","DOI":"10.1007\/s11263-016-0981-7","article-title":"Visual genome","volume":"123","author":"Krishna","year":"2017","journal-title":"Computer vision"},{"key":"2024072520071813700_ref18","doi-asserted-by":"crossref","first-page":"180251","DOI":"10.1038\/sdata.2018.251","article-title":"A dataset of clinically generated visual questions and answers about radiology images","volume":"5","author":"Lau","year":"2018","journal-title":"Scientific Data"},{"key":"2024072520071813700_ref19","first-page":"3","article-title":"Psychology of the image [in Russian]","author":"Leontiev","year":"1979","journal-title":"Vestn. Mosk. un-ta. Ser. 14, Psychology"},{"key":"2024072520071813700_ref20","first-page":"236","volume-title":"Bulletin of the YUrGGPU","author":"Lesnichaya","year":"2008"},{"key":"2024072520071813700_ref21","first-page":"2022","volume-title":"IEEE\/CVF International Conference on Computer Vision (ICCV)","author":"Li","year":"2021"},{"key":"2024072520071813700_ref22","first-page":"3190","volume-title":"IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Marino","year":"2019"},{"key":"2024072520071813700_ref23","doi-asserted-by":"crossref","first-page":"3","DOI":"10.1007\/978-3-319-27060-9_1","volume-title":"Advances in Artificial Intelligence and Soft Computing","author":"Osipov","year":"2015"},{"key":"2024072520071813700_ref24","volume-title":"Les M\u2019Ecanismes Perceptifs [in French]","author":"Piaget","year":"1961"},{"key":"2024072520071813700_ref25","volume-title":"Features of Mental Development of Preschool Children [in Russian]","author":"Poddyakov","year":"1996"},{"key":"2024072520071813700_ref26","volume-title":"The Person and the Situation: Perspectives of Social Psy- Chology","author":"Ross","year":"2011"},{"key":"2024072520071813700_ref27","volume-title":"Scripts, Plans, Goals and Understanding: An Inquiry into Human Knowledge Structures","author":"Schank","year":"1977"},{"key":"2024072520071813700_ref28","doi-asserted-by":"crossref","first-page":"146","DOI":"10.1007\/978-3-031-20074-8_9","volume-title":"Computer Vision \u2013 ECCV 2022. ECCV 2022","author":"Schwenk","year":"2022"},{"key":"2024072520071813700_ref29","volume-title":"Preschool and Primary School Age","author":"Semago","year":"2005"},{"key":"2024072520071813700_ref30","first-page":"240","article-title":"Description of the image structure in modern art criticism analysis [in Russian]","volume":"13","author":"Shapoval","year":"2011","journal-title":"Izvestiya Samarskogo nauchnogo tsentra Rossiyskoy akademii nauk"},{"key":"2024072520071813700_ref31","first-page":"20346","volume-title":"Advances in Neural Information Processing Systems 34 (NeurIPS 2021)","author":"Sheng","year":"2021"},{"key":"2024072520071813700_ref32","volume-title":"Hand-Drawn Apperceptive Test RAT","author":"Sobchik","year":"2002"},{"key":"2024072520071813700_ref33","first-page":"2542","volume-title":"IEEE International Conference on Computer Vision (ICCV)","author":"Vedantam","year":"2015"},{"key":"2024072520071813700_ref34","author":"Velichkovsky","year":"2006"},{"key":"2024072520071813700_ref35","volume-title":"Thinking and Speaking","author":"Vygotsky","year":"1962"},{"key":"2024072520071813700_ref36","first-page":"19175","volume-title":"IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Wang","year":"2023"},{"key":"2024072520071813700_ref37","doi-asserted-by":"crossref","first-page":"2413","DOI":"10.1109\/TPAMI.2017.2754246","article-title":"Fvqa: fact-based visual question answering","volume":"40","author":"Wang","year":"2018","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"2024072520071813700_ref38","doi-asserted-by":"crossref","first-page":"1290","DOI":"10.24963\/ijcai.2017\/179","volume-title":"Proceedings of the Twenty-Sixth International Joint Conference on Artificial Intelligence","author":"Wang","year":"2017"},{"key":"2024072520071813700_ref39","first-page":"23318","volume-title":"Proceedings of the 39 the International Conference on Machine Learning","author":"Wang","year":"2022"},{"key":"2024072520071813700_ref40","first-page":"6713","volume-title":"IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Zellers","year":"2019"}],"container-title":["Logic Journal of the IGPL"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/academic.oup.com\/jigpal\/article-pdf\/32\/4\/656\/58646565\/jzae026.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/academic.oup.com\/jigpal\/article-pdf\/32\/4\/656\/58646565\/jzae026.pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,7,25]],"date-time":"2024-07-25T20:09:53Z","timestamp":1721938193000},"score":1,"resource":{"primary":{"URL":"https:\/\/academic.oup.com\/jigpal\/article\/32\/4\/656\/7632103"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,3,22]]},"references-count":40,"journal-issue":{"issue":"4","published-online":{"date-parts":[[2024,3,22]]},"published-print":{"date-parts":[[2024,7,25]]}},"URL":"https:\/\/doi.org\/10.1093\/jigpal\/jzae026","relation":{},"ISSN":["1367-0751","1368-9894"],"issn-type":[{"value":"1367-0751","type":"print"},{"value":"1368-9894","type":"electronic"}],"subject":[],"published-other":{"date-parts":[[2024,8]]},"published":{"date-parts":[[2024,3,22]]}}}