{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,10]],"date-time":"2026-04-10T14:42:50Z","timestamp":1775832170993,"version":"3.50.1"},"reference-count":53,"publisher":"Springer Science and Business Media LLC","issue":"1","license":[{"start":{"date-parts":[[2023,5,5]],"date-time":"2023-05-05T00:00:00Z","timestamp":1683244800000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2023,5,5]],"date-time":"2023-05-05T00:00:00Z","timestamp":1683244800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["61932003"],"award-info":[{"award-number":["61932003"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]},{"name":"the Fundamental Research Funds for the Central Universities"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Vis. Comput. Ind. Biomed. Art"],"abstract":"<jats:title>Abstract<\/jats:title><jats:p>In recent years, deep learning techniques have been used to estimate gaze\u2014a significant task in computer vision and human-computer interaction. Previous studies have made significant achievements in predicting 2D or 3D gazes from monocular face images. This study presents a deep neural network for 2D gaze estimation on mobile devices. It achieves state-of-the-art 2D gaze point regression error, while significantly improving gaze classification error on quadrant divisions of the display. To this end, an efficient attention-based module that correlates and fuses the left and right eye contextual features is first proposed to improve gaze point regression performance. Subsequently, through a unified perspective for gaze estimation, metric learning for gaze classification on quadrant divisions is incorporated as additional supervision. Consequently, both gaze point regression and quadrant classification performances are improved. The experiments demonstrate that the proposed method outperforms existing gaze-estimation methods on the GazeCapture and MPIIFaceGaze datasets.<\/jats:p>","DOI":"10.1186\/s42492-023-00135-6","type":"journal-article","created":{"date-parts":[[2023,5,5]],"date-time":"2023-05-05T04:01:44Z","timestamp":1683259304000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":7,"title":["EM-Gaze: eye context correlation and metric learning for gaze estimation"],"prefix":"10.1186","volume":"6","author":[{"given":"Jinchao","family":"Zhou","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Guoan","family":"Li","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Feng","family":"Shi","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xiaoyan","family":"Guo","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Pengfei","family":"Wan","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Miao","family":"Wang","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2023,5,5]]},"reference":[{"key":"135_CR1","doi-asserted-by":"publisher","unstructured":"Zhang XC, Sugano Y, Fritz M, Bulling A (2015) Appearance-based gaze estimation in the wild. In: Proceedings of the 2015 IEEE conference on computer vision and pattern recognition, IEEE, Boston, 7\u201312 June 2015. https:\/\/doi.org\/10.1109\/CVPR.2015.7299081","DOI":"10.1109\/CVPR.2015.7299081"},{"key":"135_CR2","doi-asserted-by":"publisher","unstructured":"Krafka K, Khosla A, Kellnhofer P, Kannan H, Bhandarkar S, Matusik W et al (2016) Eye tracking for everyone. In: Proceedings of the 2016 IEEE conference on computer vision and pattern recognition, IEEE, Las Vegas, 27\u201330 June 2016. https:\/\/doi.org\/10.1109\/CVPR.2016.239","DOI":"10.1109\/CVPR.2016.239"},{"key":"135_CR3","doi-asserted-by":"publisher","unstructured":"He JF, Pham K, Valliappan N, Xu PM, Roberts C, Lagun D et al (2019) On-device few-shot personalization for real-time gaze estimation. In: Proceedings of the 2019 IEEE\/CVF international conference on computer vision workshop, IEEE, Seoul, 27\u201328 October 2019. https:\/\/doi.org\/10.1109\/ICCVW.2019.00146","DOI":"10.1109\/ICCVW.2019.00146"},{"key":"135_CR4","doi-asserted-by":"publisher","unstructured":"Bao YW, Cheng YH, Liu YF, Lu F (2021) Adaptive feature fusion network for gaze tracking in mobile tablets. In: Proceedings of the 2020 25th international conference on pattern recognition, IEEE, Milan, 10\u201315 January 2021. https:\/\/doi.org\/10.1109\/ICPR48806.2021.9412205","DOI":"10.1109\/ICPR48806.2021.9412205"},{"issue":"1","key":"135_CR5","doi-asserted-by":"publisher","first-page":"24","DOI":"10.1186\/s42492-019-0034-5","volume":"2","author":"I Dagher","year":"2019","unstructured":"Dagher I, Dahdah E, Al Shakik M (2019) Facial expression recognition using three-stage support vector machines. Vis Comput Ind Biomed Art 2(1):24. https:\/\/doi.org\/10.1186\/s42492-019-0034-5","journal-title":"Vis Comput Ind Biomed Art"},{"key":"135_CR6","doi-asserted-by":"publisher","DOI":"10.1109\/TMM.2022.3156820","author":"SY Chen","year":"2022","unstructured":"Chen SY, Lai YK, Xia SH, Rosin P, Gao L (2022) 3D face reconstruction and gaze tracking in the HMD for virtual interaction. IEEE Trans Multimedia. https:\/\/doi.org\/10.1109\/TMM.2022.3156820","journal-title":"IEEE Trans Multimedia"},{"key":"135_CR7","doi-asserted-by":"publisher","unstructured":"Modi N, Singh J (2021) A review of various state of art eye gaze estimation techniques. In: Gao XZ, Tiwari S, Trivedi M, Mishra K (eds) Advances in computational intelligence and communication technology. Advances in intelligent systems and computing, vol. 1086. Springer, Singapore, pp 501\u2013510. https:\/\/doi.org\/10.1007\/978-981-15-1275-9_41","DOI":"10.1007\/978-981-15-1275-9_41"},{"issue":"3","key":"135_CR8","doi-asserted-by":"publisher","first-page":"478","DOI":"10.1109\/TPAMI.2009.30","volume":"32","author":"DW Hansen","year":"2010","unstructured":"Hansen DW, Ji Q (2010) In the eye of the beholder: a survey of models for eyes and gaze. IEEE Trans Pattern Anal Mach Intell 32(3):478-500. https:\/\/doi.org\/10.1109\/TPAMI.2009.30","journal-title":"IEEE Trans Pattern Anal Mach Intell"},{"key":"135_CR9","unstructured":"Cheng YH, Wang HF, Bao YW, Lu F (2021) Appearance-based gaze estimation with deep learning: a review and benchmark. arXiv: 2104.12668"},{"issue":"5","key":"135_CR10","doi-asserted-by":"publisher","first-page":"2002","DOI":"10.1109\/TVCG.2019.2899187","volume":"25","author":"ZM Hu","year":"2019","unstructured":"Hu ZM, Zhang CY, Li S, Wang GP, Manocha D (2019) SGaze: a data-driven eye-head coordination model for realtime gaze prediction. IEEE Trans Vis Comput Graph 25(5):2002-2010. https:\/\/doi.org\/10.1109\/TVCG.2019.2899187","journal-title":"IEEE Trans Vis Comput Graph"},{"issue":"5","key":"135_CR11","doi-asserted-by":"publisher","first-page":"1902","DOI":"10.1109\/TVCG.2020.2973473","volume":"26","author":"ZM Hu","year":"2020","unstructured":"Hu ZM, Li S, Zhang CY, Yi KR, Wang GP, Manocha D (2020) DGaze: CNN-based gaze prediction in dynamic scenes. IEEE Trans Vis Comput Graph 26(5):1902-1911. https:\/\/doi.org\/10.1109\/TVCG.2020.2973473","journal-title":"IEEE Trans Vis Comput Graph"},{"issue":"6","key":"135_CR12","doi-asserted-by":"publisher","first-page":"1124","DOI":"10.1109\/TBME.2005.863952","volume":"53","author":"ED Guestrin","year":"2006","unstructured":"Guestrin ED, Eizenman M (2006) General theory of remote gaze estimation using the pupil center and corneal reflections. IEEE Trans Biomed Eng 53(6):1124-1133. https:\/\/doi.org\/10.1109\/TBME.2005.863952","journal-title":"IEEE Trans Biomed Eng"},{"key":"135_CR13","doi-asserted-by":"publisher","unstructured":"Nakazawa A, Nitschke C (2012) Point of gaze estimation through corneal surface reflection in an active illumination environment. In: Fitzgibbon A, Lazebnik S, Perona P, Sato Y, Schmid C (eds) Computer vision - ECCV 2012. 12th European conference on computer vision, Florence, Italy, October 7\u201313, 2012. Lecture notes in computer science, vol. 7573. Springer, Florence, pp 159\u2013172. https:\/\/doi.org\/10.1007\/978-3-642-33709-3_12","DOI":"10.1007\/978-3-642-33709-3_12"},{"issue":"2","key":"135_CR14","doi-asserted-by":"publisher","first-page":"802","DOI":"10.1109\/TIP.2011.2162740","volume":"21","author":"R Valenti","year":"2012","unstructured":"Valenti R, Sebe N, Gevers T (2012) Combining head pose and eye location information for gaze estimation. IEEE Trans Image Process 21(2):802-815. https:\/\/doi.org\/10.1109\/TIP.2011.2162740","journal-title":"IEEE Trans Image Process"},{"key":"135_CR15","doi-asserted-by":"publisher","unstructured":"Alberto Funes Mora K, Odobez JM (2014) Geometric generative gaze estimation (G3E) for remote RGB-d cameras. In: Proceedings of the 2014 IEEE conference on computer vision and pattern recognition, IEEE, Columbus, 23\u201328 June 2014. https:\/\/doi.org\/10.1109\/CVPR.2014.229","DOI":"10.1109\/CVPR.2014.229"},{"key":"135_CR16","doi-asserted-by":"publisher","unstructured":"Xiong XH, Liu ZC, Cai Q, Zhang ZY (2014) Eye gaze tracking using an RGBD camera: a comparison with a RGB solution. In: Proceedings of the 2014 ACM international joint conference on pervasive and ubiquitous computing: adjunct publication, ACM, Seattle, 13 September 2014. https:\/\/doi.org\/10.1145\/2638728.2641694","DOI":"10.1145\/2638728.2641694"},{"issue":"3","key":"135_CR17","doi-asserted-by":"publisher","first-page":"543","DOI":"10.1007\/s11042-012-1202-1","volume":"65","author":"YT Lin","year":"2013","unstructured":"Lin YT, Lin RY, Lin YC, Lee GC (2013) Real-time eye-gaze estimation using a low-resolution webcam. Multimed Tools Appl 65(3):543-568. https:\/\/doi.org\/10.1007\/s11042-012-1202-1","journal-title":"Multimed Tools Appl"},{"issue":"10","key":"135_CR18","doi-asserted-by":"publisher","first-page":"2033","DOI":"10.1109\/TPAMI.2014.2313123","volume":"36","author":"F Lu","year":"2014","unstructured":"Lu F, Sugano Y, Okabe T, Sato Y (2014) Adaptive linear regression for appearance-based gaze estimation. IEEE Trans Pattern Anal Mach Intell 36(10):2033-2046. https:\/\/doi.org\/10.1109\/TPAMI.2014.2313123","journal-title":"IEEE Trans Pattern Anal Mach Intell"},{"key":"135_CR19","unstructured":"Williams O, Blake A, Cipolla R (2006) Sparse and semi-supervised visual mapping with the S3GP. In: Proceedings of the 2006 IEEE computer society conference on computer vision and pattern recognition, IEEE, New York, 17\u201322 June 2006"},{"issue":"4","key":"135_CR20","doi-asserted-by":"publisher","first-page":"1543","DOI":"10.1109\/TIP.2017.2657880","volume":"26","author":"F Lu","year":"2017","unstructured":"Lu F, Chen XW, Sato Y (2017) Appearance-based gaze estimation via uncalibrated gaze pattern recovery. IEEE Trans Image Process 26(4):1543-1553. https:\/\/doi.org\/10.1109\/TIP.2017.2657880","journal-title":"IEEE Trans Image Process"},{"issue":"11","key":"135_CR21","doi-asserted-by":"publisher","first-page":"2278","DOI":"10.1109\/5.726791","volume":"86","author":"Y LeCun","year":"1998","unstructured":"LeCun Y, Bottou L, Bengio Y, Haffner P (1998) Gradient-based learning applied to document recognition. Proc IEEE 86(11):2278-2324. https:\/\/doi.org\/10.1109\/5.726791","journal-title":"Proc IEEE"},{"key":"135_CR22","doi-asserted-by":"publisher","unstructured":"Yu Y, Liu G, Odobez JM (2018) Deep multitask gaze estimation with a constrained landmark-gaze model. In: Leal-Taix\u00e9 L, Roth S (eds) Computer vision - ECCV 2018 workshops. Munich, Germany, September 8\u201314, 2018, Proceedings, Part II. Lecture notes in computer science, vol. 11130. Springer, Munich, pp 456\u2013474. https:\/\/doi.org\/10.1007\/978-3-030-11012-3_35","DOI":"10.1007\/978-3-030-11012-3_35"},{"key":"135_CR23","doi-asserted-by":"publisher","unstructured":"Fischer T, Chang HJ, Demiris Y (2018) RT-GENE: real-time eye gaze estimation in natural environments. In: Ferrari V, Hebert M, Sminchisescu C, Weiss Y (eds) Computer vision - ECCV 2018. 15th European conference, Munich, Germany, September 8\u201314, 2018. Lecture notes in computer science. Springer, Munich, pp 339\u2013357. https:\/\/doi.org\/10.1007\/978-3-030-01249-6_21","DOI":"10.1007\/978-3-030-01249-6_21"},{"key":"135_CR24","unstructured":"Simonyan K, Zisserman A (2014) Very deep convolutional networks for large-scale image recognition. arXiv preprint arXiv:1409.1556"},{"key":"135_CR25","doi-asserted-by":"publisher","unstructured":"Cheng YH, Lu F, Zhang XC (2018) Appearance-based gaze estimation via evaluation-guided asymmetric regression. In: Ferrari V, Hebert M, Sminchisescu C, Weiss Y (eds) Computer vision - ECCV 2018. 15th European conference, Munich, Germany, September 8\u201314, 2018. Lecture notes in computer science, vol. 11218. Springer, Munich, pp 105\u2013121. https:\/\/doi.org\/10.1007\/978-3-030-01264-9_7","DOI":"10.1007\/978-3-030-01264-9_7"},{"key":"135_CR26","doi-asserted-by":"publisher","unstructured":"Park S, de Mello S, Molchanov P, Iqbal U, Hilliges O, Kautz J (2019) Few-shot adaptive gaze estimation. In: Proceedings of the 2019 IEEE\/CVF international conference on computer vision, IEEE, Seoul, 27 October\u20132 November 2019. https:\/\/doi.org\/10.1109\/ICCV.2019.00946","DOI":"10.1109\/ICCV.2019.00946"},{"key":"135_CR27","doi-asserted-by":"publisher","unstructured":"Cheng YH, Lu F (2022) Gaze estimation using transformer. In: Proceedings of the 2022 26th international conference on pattern recognition, IEEE, Montreal, 21\u201325 August 2022. https:\/\/doi.org\/10.1109\/ICPR56361.2022.9956687","DOI":"10.1109\/ICPR56361.2022.9956687"},{"key":"135_CR28","doi-asserted-by":"publisher","unstructured":"Cheng YH, Bao YW, Lu F (2022) Puregaze: purifying gaze feature for generalizable gaze estimation. In: Proceedings of the 36th AAAI conference on artificial intelligence, AAAI Press, Vancouver, 22 February-1 March 2022. https:\/\/doi.org\/10.1609\/aaai.v36i1.19921","DOI":"10.1609\/aaai.v36i1.19921"},{"issue":"2","key":"135_CR29","doi-asserted-by":"publisher","first-page":"179","DOI":"10.1109\/TCE.2019.2899869","volume":"65","author":"J Lemley","year":"2019","unstructured":"Lemley J, Kar A, Drimbarean A, Corcoran P (2019) Convolutional neural network implementation for eye-gaze estimation on low-quality consumer imaging systems. IEEE Trans Consum Electron 65(2):179-187. https:\/\/doi.org\/10.1109\/TCE.2019.2899869","journal-title":"IEEE Trans Consum Electron"},{"issue":"4","key":"135_CR30","doi-asserted-by":"publisher","first-page":"166","DOI":"10.1145\/3528223.3530130","volume":"41","author":"GY Li","year":"2022","unstructured":"Li GY, Meka A, Mueller F, Buehler MC, Hilliges O, Beeler T (2022) EyeNeRF: a hybrid representation for photorealistic synthesis, animation and relighting of human eyes. ACM Trans Graph 41(4):166. https:\/\/doi.org\/10.1145\/3528223.3530130","journal-title":"ACM Trans Graph"},{"key":"135_CR31","doi-asserted-by":"publisher","unstructured":"Zhang XC, Sugano Y, Fritz M, Bulling A (2017) It\u2019s written all over your face: full-face appearance-based gaze estimation. In: Proceedings of the 2017 IEEE conference on computer vision and pattern recognition workshops. IEEE, Honolulu, 21\u201326 July 2017. https:\/\/doi.org\/10.1109\/CVPRW.2017.284","DOI":"10.1109\/CVPRW.2017.284"},{"key":"135_CR32","doi-asserted-by":"publisher","unstructured":"Schroff F, Kalenichenko D, Philbin J (2015) FaceNet: a unified embedding for face recognition and clustering. In: Proceedings of the 2015 IEEE conference on computer vision and pattern recognition, IEEE, Boston, 7\u201312 June 2015. https:\/\/doi.org\/10.1109\/CVPR.2015.7298682","DOI":"10.1109\/CVPR.2015.7298682"},{"key":"135_CR33","unstructured":"Hermans A, Beyer L, Leibe B (2017) In defense of the triplet loss for person re-identification. arXiv: 1703.07737"},{"key":"135_CR34","unstructured":"Liu WY, Wen YD, Yu ZD, Yang M (2016) Large-margin softmax loss for convolutional neural networks. In: Proceedings of the 33rd international conference on international conference on machine learning, JMLR.org, New York, 19 June 2016"},{"key":"135_CR35","doi-asserted-by":"publisher","unstructured":"Liu WY, Wen YD, Yu ZD, Li M, Raj B, Song L (2017) SphereFace: deep hypersphere embedding for face recognition. In: Proceedings of the 2017 IEEE conference on computer vision and pattern recognition, IEEE, Honolulu, 21\u201326 July 2017. https:\/\/doi.org\/10.1109\/CVPR.2017.713","DOI":"10.1109\/CVPR.2017.713"},{"key":"135_CR36","doi-asserted-by":"publisher","unstructured":"Wang H, Wang YT, Zhou Z, Ji X, Gong DH, Zhou JC et al (2018) CosFace: large margin cosine loss for deep face recognition. In: Proceedings of the 2018 IEEE\/CVF conference on computer vision and pattern recognition, IEEE, Salt Lake City, 18\u201323 June 2018. https:\/\/doi.org\/10.1109\/CVPR.2018.00552","DOI":"10.1109\/CVPR.2018.00552"},{"key":"135_CR37","doi-asserted-by":"publisher","unstructured":"Musgrave K, Belongie S, Lim SN (2020) A metric learning reality check. In: Vedaldi A, Bischof H, Brox T, Frahm JM (eds) Computer vision - ECCV 2020. 16th European conference, Glasgow, UK, August 23\u201328, 2020. Lecture notes in computer science, vol. 12370. Springer, Glasgow, pp 681\u2013699. https:\/\/doi.org\/10.1007\/978-3-030-58595-2_41","DOI":"10.1007\/978-3-030-58595-2_41"},{"key":"135_CR38","doi-asserted-by":"publisher","unstructured":"Hu J, Shen L, Sun G (2018) Squeeze-and-excitation networks. In: Proceedings of the 2018 IEEE\/CVF conference on computer vision and pattern recognition, IEEE, Salt Lake City, 18\u201323 June 2018. https:\/\/doi.org\/10.1109\/CVPR.2018.00745","DOI":"10.1109\/CVPR.2018.00745"},{"key":"135_CR39","unstructured":"Vaswani A, Shazeer N, Parmar N, Uszkoreit J, Jones L, Gomez AN et al (2017) Attention is all you need. In: Proceedings of the 31st international conference on neural information processing systems, Curran Associates Inc., Long Beach, 4 December 2017"},{"key":"135_CR40","unstructured":"Dosovitskiy A, Beyer L, Kolesnikov A, Weissenborn D, Zhai XH, Unterthiner T et al (2020) An image is worth 16x16 words: transformers for image recognition at scale. In: Proceedings of the 9th international conference on learning representations, OpenReview.net, 3\u20137 May 2021"},{"issue":"2","key":"135_CR41","doi-asserted-by":"publisher","first-page":"1489","DOI":"10.1109\/TPAMI.2022.3164083","volume":"45","author":"YH Li","year":"2023","unstructured":"Li YH, Yao T, Pan YW, Mei T (2023) Contextual transformer networks for visual recognition. IEEE Trans Pattern Anal Mach Intell 45(2):1489-1500. https:\/\/doi.org\/10.1109\/TPAMI.2022.3164083","journal-title":"IEEE Trans Pattern Anal Mach Intell"},{"key":"135_CR42","doi-asserted-by":"publisher","unstructured":"Chen CFR, Fan QF, Panda R (2021) CrossViT: cross-attention multi-scale vision transformer for image classification. In: Proceedings of the 2021 IEEE\/CVF international conference on computer vision, IEEE, Montreal, 10\u201317 October 2021. https:\/\/doi.org\/10.1109\/ICCV48922.2021.00041","DOI":"10.1109\/ICCV48922.2021.00041"},{"issue":"1","key":"135_CR43","doi-asserted-by":"publisher","first-page":"7","DOI":"10.1186\/s42492-020-00045-x","volume":"3","author":"L Chen","year":"2020","unstructured":"Chen L, Liu R, Zhou DS, Yang X, Zhang Q (2020) Fused behavior recognition model based on attention mechanism. Vis Comput Ind Biomed Art 3(1):7. https:\/\/doi.org\/10.1186\/s42492-020-00045-x","journal-title":"Vis Comput Ind Biomed Art"},{"issue":"1","key":"135_CR44","doi-asserted-by":"publisher","first-page":"12","DOI":"10.1186\/s42492-022-00110-7","volume":"5","author":"WW Yuan","year":"2022","unstructured":"Yuan WW, Peng YJ, Guo YF, Ren YD, Xue QW (2022) Correction: DCAU-Net: dense convolutional attention u-net for segmentation of intracranial aneurysm images. Vis Comput Ind Biomed Art 5(1):12. https:\/\/doi.org\/10.1186\/s42492-022-00110-7","journal-title":"Vis Comput Ind Biomed Art"},{"key":"135_CR45","doi-asserted-by":"publisher","unstructured":"Cheng YH, Huang SY, Wang F, Qian C, Lu F (2020) A coarse-to-fine adaptive network for appearance-based gaze estimation. In: Proceedings of the 34th AAAI conference on artificial intelligence, AAAI Press, New York, 7\u201312 February 2020. https:\/\/doi.org\/10.1609\/aaai.v34i07.6636","DOI":"10.1609\/aaai.v34i07.6636"},{"key":"135_CR46","doi-asserted-by":"publisher","unstructured":"Wang XL, Girshick R, Gupta A, He KM (2018) Non-local neural networks. In: Proceedings of the 2018 IEEE\/CVF conference on computer vision and pattern recognition, IEEE, Salt Lake City, 18\u201323 June 2018. https:\/\/doi.org\/10.1109\/CVPR.2018.00813","DOI":"10.1109\/CVPR.2018.00813"},{"key":"135_CR47","doi-asserted-by":"publisher","unstructured":"Li X, Wang WH, Hu XL, Yang J (2019) Selective kernel networks. In: Proceedings of the 2019 IEEE\/CVF conference on computer vision and pattern recognition, IEEE, Long Beach, 15\u201320 June 2019. https:\/\/doi.org\/10.1109\/CVPR.2019.00060","DOI":"10.1109\/CVPR.2019.00060"},{"key":"135_CR48","unstructured":"Tolstikhin IO, Houlsby N, Kolesnikov A, Beyer L, Zhai XH, Unterthiner T et al (2021) MLP-mixer: an all-MLP architecture for vision. In: Proceedings of Advances in Neural Information Processing Systems 34 (NeurIPS 2021), online, 6\u201314 December 2021"},{"key":"135_CR49","volume-title":"Deep learning","author":"I Goodfellow","year":"2016","unstructured":"Goodfellow I, Bengio Y, Courville A (2016) Deep learning. MIT Press, Cambridge"},{"key":"135_CR50","unstructured":"Glorot X, Bengio Y (2010) Understanding the difficulty of training deep feedforward neural networks. In: Proceedings of the 13th international conference on artificial intelligence and statistics, JMLR.org, Sardinia, 13\u201315 May 2010"},{"key":"135_CR51","unstructured":"Paszke A, Gross S, Massa F, Lerer A, Bradbury J, Chanan G et al (2019) PyTorch: an imperative style, high-performance deep learning library. In: Proceedings of the 33rd conference on neural information processing systems. Curran Associates Inc., Vancouver, 8 December 2019"},{"key":"135_CR52","doi-asserted-by":"publisher","unstructured":"Guo TC, Liu YC, Zhang H, Liu XB, Kwak Y, In Yoo B et al (2019) A generalized and robust method towards practical gaze estimation on smart phone. In: Proceedings of the 2019 IEEE\/CVF international conference on computer vision workshop, IEEE, Seoul, 27\u201328 October 2019. https:\/\/doi.org\/10.1109\/ICCVW.2019.00144","DOI":"10.1109\/ICCVW.2019.00144"},{"issue":"86","key":"135_CR53","first-page":"2579","volume":"9","author":"L van der Maaten","year":"2008","unstructured":"van der Maaten L, Hinton G (2008) Visualizing data using t-SNE. J Mach Learn Res 9(86):2579-2605","journal-title":"J Mach Learn Res"}],"container-title":["Visual Computing for Industry, Biomedicine, and Art"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1186\/s42492-023-00135-6.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1186\/s42492-023-00135-6\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1186\/s42492-023-00135-6.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2023,5,5]],"date-time":"2023-05-05T04:02:55Z","timestamp":1683259375000},"score":1,"resource":{"primary":{"URL":"https:\/\/vciba.springeropen.com\/articles\/10.1186\/s42492-023-00135-6"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,5,5]]},"references-count":53,"journal-issue":{"issue":"1","published-online":{"date-parts":[[2023,12]]}},"alternative-id":["135"],"URL":"https:\/\/doi.org\/10.1186\/s42492-023-00135-6","relation":{},"ISSN":["2524-4442"],"issn-type":[{"value":"2524-4442","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,5,5]]},"assertion":[{"value":"18 November 2022","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"15 April 2023","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"5 May 2023","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"The authors declare no competing interests.","order":2,"name":"Ethics","group":{"name":"EthicsHeading","label":"Competing interests"}}],"article-number":"8"}}