{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,3,16]],"date-time":"2026-03-16T21:15:32Z","timestamp":1773695732553,"version":"3.50.1"},"publisher-location":"New York, NY, USA","reference-count":63,"publisher":"ACM","license":[{"start":{"date-parts":[[2021,8,24]],"date-time":"2021-08-24T00:00:00Z","timestamp":1629763200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"University of Amsterdam"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2021,8,24]]},"DOI":"10.1145\/3460426.3463667","type":"proceedings-article","created":{"date-parts":[[2021,9,1]],"date-time":"2021-09-01T22:50:29Z","timestamp":1630536629000},"page":"645-652","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":21,"title":["Contextualized Keyword Representations for Multi-modal Retinal Image Captioning"],"prefix":"10.1145","author":[{"given":"Jia-Hong","family":"Huang","sequence":"first","affiliation":[{"name":"University of Amsterdam, Amsterdam, Netherlands"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Ting-Wei","family":"Wu","sequence":"additional","affiliation":[{"name":"Georgia Institute of Technology, Atlanta, GA, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Marcel","family":"Worring","sequence":"additional","affiliation":[{"name":"University of Amsterdam, Amsterdam, Netherlands"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2021,9]]},"reference":[{"key":"e_1_3_2_1_1_1","volume-title":"Accuracy assessment of intra-and intervisit fundus image registration for diabetic retinopathy screening. Investigative ophthalmology & visual science","author":"Adal Kedir M","year":"2015","unstructured":"Kedir M Adal , Peter G van Etten , Jose P Martinez , Lucas J van Vliet , and Koenraad A Vermeer . 2015. Accuracy assessment of intra-and intervisit fundus image registration for diabetic retinopathy screening. Investigative ophthalmology & visual science , Vol. 56 , 3 ( 2015 ), 1805--1812. Kedir M Adal, Peter G van Etten, Jose P Martinez, Lucas J van Vliet, and Koenraad A Vermeer. 2015. Accuracy assessment of intra-and intervisit fundus image registration for diabetic retinopathy screening. Investigative ophthalmology & visual science , Vol. 56, 3 (2015), 1805--1812."},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-016-0966-6"},{"key":"e_1_3_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1109\/IEMBS.2008.4649647"},{"key":"e_1_3_2_1_4_1","volume-title":"Proceedings of the acl workshop on intrinsic and extrinsic evaluation measures for machine translation and\/or summarization . 65--72","author":"Banerjee Satanjeev","year":"2005","unstructured":"Satanjeev Banerjee and Alon Lavie . 2005 . METEOR: An automatic metric for MT evaluation with improved correlation with human judgments . In Proceedings of the acl workshop on intrinsic and extrinsic evaluation measures for machine translation and\/or summarization . 65--72 . Satanjeev Banerjee and Alon Lavie. 2005. METEOR: An automatic metric for MT evaluation with improved correlation with human judgments. In Proceedings of the acl workshop on intrinsic and extrinsic evaluation measures for machine translation and\/or summarization . 65--72."},{"key":"e_1_3_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.artmed.2008.04.005"},{"key":"e_1_3_2_1_6_1","volume-title":"United Kingdom","author":"Computing Retinal Image","year":"2012","unstructured":"Retinal Image Computing . 2012 . Understanding,?ONHSD-Optic Nerve Head Segmentation Dataset,\" University of Lincoln , United Kingdom , 2004. Retinal Image Computing. 2012. Understanding,?ONHSD-Optic Nerve Head Segmentation Dataset,\" University of Lincoln, United Kingdom, 2004."},{"key":"e_1_3_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.5566\/ias.1155"},{"key":"e_1_3_2_1_8_1","volume-title":"Bert: Pre-training of deep bidirectional transformers for language understanding. arXiv preprint arXiv:1810.04805","author":"Devlin Jacob","year":"2018","unstructured":"Jacob Devlin , Ming-Wei Chang , Kenton Lee , and Kristina Toutanova . 2018 . Bert: Pre-training of deep bidirectional transformers for language understanding. arXiv preprint arXiv:1810.04805 (2018). Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018. Bert: Pre-training of deep bidirectional transformers for language understanding. arXiv preprint arXiv:1810.04805 (2018)."},{"key":"e_1_3_2_1_9_1","volume-title":"How contextual are contextualized word representations? comparing the geometry of BERT, ELMo, and GPT-2 embeddings. arXiv preprint arXiv:1909.00512","author":"Ethayarajh Kawin","year":"2019","unstructured":"Kawin Ethayarajh . 2019. How contextual are contextualized word representations? comparing the geometry of BERT, ELMo, and GPT-2 embeddings. arXiv preprint arXiv:1909.00512 ( 2019 ). Kawin Ethayarajh. 2019. How contextual are contextualized word representations? comparing the geometry of BERT, ELMo, and GPT-2 embeddings. arXiv preprint arXiv:1909.00512 (2019)."},{"key":"e_1_3_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2015.7298754"},{"key":"e_1_3_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/TBME.2012.2205687"},{"key":"e_1_3_2_1_12_1","volume-title":"Deliberate Attention Networks for Image Captioning. AAAI","author":"Gao Lianli","year":"2019","unstructured":"Lianli Gao , Kaixuan Fan , Jingkuan Song , Xianglong Liu , Xing Xu , and Heng Tao Shen . 2019. Deliberate Attention Networks for Image Captioning. AAAI ( 2019 ). Lianli Gao, Kaixuan Fan, Jingkuan Song, Xianglong Liu, Xing Xu, and Heng Tao Shen. 2019. Deliberate Attention Networks for Image Captioning. AAAI (2019)."},{"key":"e_1_3_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1080\/00437956.1954.11659520"},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-46493-0_1"},{"key":"e_1_3_2_1_15_1","first-page":"16","article-title":"FIRE: fundus image registration dataset","volume":"1","author":"Hernandez-Matas Carlos","year":"2017","unstructured":"Carlos Hernandez-Matas , Xenophon Zabulis , Areti Triantafyllou , Panagiota Anyfanti , Stella Douma , and Antonis A Argyros . 2017 . FIRE: fundus image registration dataset . Journal for Modeling in Ophthalmology , Vol. 1 , 4 (2017), 16 -- 28 . Carlos Hernandez-Matas, Xenophon Zabulis, Areti Triantafyllou, Panagiota Anyfanti, Stella Douma, and Antonis A Argyros. 2017. FIRE: fundus image registration dataset. Journal for Modeling in Ophthalmology , Vol. 1, 4 (2017), 16--28.","journal-title":"Journal for Modeling in Ophthalmology"},{"key":"e_1_3_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/TMI.2003.815900"},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00517"},{"key":"e_1_3_2_1_18_1","volume-title":"Robustness Analysis of Visual Question Answering Models by Basic Questions","author":"Huang Jia-Hong","year":"2017","unstructured":"Jia-Hong Huang . 2017. Robustness Analysis of Visual Question Answering Models by Basic Questions . King Abdullah University of Science and Technology MS thesis ( 2017 ). Jia-Hong Huang. 2017. Robustness Analysis of Visual Question Answering Models by Basic Questions. King Abdullah University of Science and Technology MS thesis (2017)."},{"key":"e_1_3_2_1_19_1","volume-title":"CVPR VQA Challenge Workshop","author":"Huang Jia-Hong","year":"2017","unstructured":"Jia-Hong Huang , Modar Alfadly , and Bernard Ghanem . 2017 . Vqabq: Visual question answering by basic questions . CVPR VQA Challenge Workshop (2017). Jia-Hong Huang, Modar Alfadly, and Bernard Ghanem. 2017. Vqabq: Visual question answering by basic questions. CVPR VQA Challenge Workshop (2017)."},{"key":"e_1_3_2_1_20_1","volume-title":"2019 a. Assessing the robustness of visual question answering. arXiv preprint arXiv:1912.01452","author":"Huang Jia-Hong","year":"2019","unstructured":"Jia-Hong Huang , Modar Alfadly , Bernard Ghanem , and Marcel Worring . 2019 a. Assessing the robustness of visual question answering. arXiv preprint arXiv:1912.01452 ( 2019 ). Jia-Hong Huang, Modar Alfadly, Bernard Ghanem, and Marcel Worring. 2019 a. Assessing the robustness of visual question answering. arXiv preprint arXiv:1912.01452 (2019)."},{"key":"e_1_3_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v33i01.33018449"},{"key":"e_1_3_2_1_22_1","volume-title":"CVPR VQA Challenge and Visual Dialog Workshop","author":"Huang Jia-Hong","year":"2018","unstructured":"Jia-Hong Huang , Cuong Duc Dao , Modar Alfadly , C Huck Yang , and Bernard Ghanem . 2018 . Robustness analysis of visual qa models by basic questions . CVPR VQA Challenge and Visual Dialog Workshop (2018). Jia-Hong Huang, Cuong Duc Dao, Modar Alfadly, C Huck Yang, and Bernard Ghanem. 2018. Robustness analysis of visual qa models by basic questions. CVPR VQA Challenge and Visual Dialog Workshop (2018)."},{"key":"e_1_3_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1145\/3460426.3463662"},{"key":"e_1_3_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1145\/3372278.3390695"},{"key":"e_1_3_2_1_25_1","volume-title":"Deep Context-Encoding Network for Retinal Image Captioning. IEEE International Conference on Image Processing (ICIP) .","author":"Huang Jia-Hong","year":"2021","unstructured":"Jia-Hong Huang , Ting-Wei Wu , Chao-Han Huck Yang , and Marcel Worring . 2021 b . Deep Context-Encoding Network for Retinal Image Captioning. IEEE International Conference on Image Processing (ICIP) . Jia-Hong Huang, Ting-Wei Wu, Chao-Han Huck Yang, and Marcel Worring. 2021 b. Deep Context-Encoding Network for Retinal Image Captioning. IEEE International Conference on Image Processing (ICIP) ."},{"key":"e_1_3_2_1_26_1","doi-asserted-by":"crossref","unstructured":"Jia-Hong Huang Ting-Wei Wu Chao-Han Huck Yang and Marcel Worring. 2021 c. Longer Version for \u201cDeep Context-Encoding Network for Retinal Image Captioning\u201d. arXiv preprint arXiv:2105.14538 .  Jia-Hong Huang Ting-Wei Wu Chao-Han Huck Yang and Marcel Worring. 2021 c. Longer Version for \u201cDeep Context-Encoding Network for Retinal Image Captioning\u201d. arXiv preprint arXiv:2105.14538 .","DOI":"10.1109\/ICIP42928.2021.9506803"},{"key":"e_1_3_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/WACV48630.2021.00249"},{"key":"e_1_3_2_1_28_1","volume-title":"On the automatic generation of medical imaging reports. ACL","author":"Jing Baoyu","year":"2018","unstructured":"Baoyu Jing , Pengtao Xie , Eric Xing , Baoyu Jing , Pengtao Xie , and Eric Xing . 2018. On the automatic generation of medical imaging reports. ACL ( 2018 ). Baoyu Jing, Pengtao Xie, Eric Xing, Baoyu Jing, Pengtao Xie, and Eric Xing. 2018. On the automatic generation of medical imaging reports. ACL (2018)."},{"key":"e_1_3_2_1_29_1","doi-asserted-by":"crossref","unstructured":"Andrej Karpathy and Li Fei-Fei. 2015. Deep visual-semantic alignments for generating image descriptions. In CVPR . 3128--3137.  Andrej Karpathy and Li Fei-Fei. 2015. Deep visual-semantic alignments for generating image descriptions. In CVPR . 3128--3137.","DOI":"10.1109\/CVPR.2015.7298932"},{"key":"e_1_3_2_1_30_1","doi-asserted-by":"crossref","unstructured":"Tomi Kauppi Valentina Kalesnykiene Joni-Kristian Kamarainen Lasse Lensu Iiris Sorri A Raninen R Voutilainen J Pietil\"a H K\"alvi\"ainen and H Uusitalo. 2007. DIARETDB1-Standard Diabetic Retinopathy Database Calibration level 1.  Tomi Kauppi Valentina Kalesnykiene Joni-Kristian Kamarainen Lasse Lensu Iiris Sorri A Raninen R Voutilainen J Pietil\"a H K\"alvi\"ainen and H Uusitalo. 2007. DIARETDB1-Standard Diabetic Retinopathy Database Calibration level 1.","DOI":"10.5244\/C.21.15"},{"key":"e_1_3_2_1_31_1","volume-title":"Adam: A method for stochastic optimization. arXiv preprint arXiv:1412.6980","author":"Kingma Diederik P","year":"2014","unstructured":"Diederik P Kingma and Jimmy Ba . 2014 . Adam: A method for stochastic optimization. arXiv preprint arXiv:1412.6980 (2014). Diederik P Kingma and Jimmy Ba. 2014. Adam: A method for stochastic optimization. arXiv preprint arXiv:1412.6980 (2014)."},{"key":"e_1_3_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-00934-2_62"},{"key":"e_1_3_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.3115\/v1\/W14-1618"},{"key":"e_1_3_2_1_34_1","first-page":"2177","article-title":"Neural word embedding as implicit matrix factorization","volume":"27","author":"Levy Omer","year":"2014","unstructured":"Omer Levy and Yoav Goldberg . 2014 b. Neural word embedding as implicit matrix factorization . NIPS , Vol. 27 (2014), 2177 -- 2185 . Omer Levy and Yoav Goldberg. 2014b. Neural word embedding as implicit matrix factorization. NIPS , Vol. 27 (2014), 2177--2185.","journal-title":"NIPS"},{"key":"e_1_3_2_1_35_1","unstructured":"Yuan Li Xiaodan Liang Zhiting Hu and Eric P Xing. 2018. Hybrid retrieval-generation reinforced agent for medical image report generation. In Advances in Neural Information Processing Systems. 1530--1540.  Yuan Li Xiaodan Liang Zhiting Hu and Eric P Xing. 2018. Hybrid retrieval-generation reinforced agent for medical image report generation. In Advances in Neural Information Processing Systems. 1530--1540."},{"key":"e_1_3_2_1_36_1","volume-title":"Rouge: A package for automatic evaluation of summaries. Text Summarization Branches Out","author":"Lin Chin-Yew","year":"2004","unstructured":"Chin-Yew Lin . 2004 . Rouge: A package for automatic evaluation of summaries. Text Summarization Branches Out (2004). Chin-Yew Lin. 2004. Rouge: A package for automatic evaluation of summaries. Text Summarization Branches Out (2004)."},{"key":"e_1_3_2_1_37_1","volume-title":"Linguistic knowledge and transferability of contextual representations. arXiv preprint arXiv:1903.08855","author":"Liu Nelson F","year":"2019","unstructured":"Nelson F Liu , Matt Gardner , Yonatan Belinkov , Matthew E Peters , and Noah A Smith . 2019. Linguistic knowledge and transferability of contextual representations. arXiv preprint arXiv:1903.08855 ( 2019 ). Nelson F Liu, Matt Gardner, Yonatan Belinkov, Matthew E Peters, and Noah A Smith. 2019. Linguistic knowledge and transferability of contextual representations. arXiv preprint arXiv:1903.08855 (2019)."},{"key":"e_1_3_2_1_38_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.100"},{"key":"e_1_3_2_1_39_1","volume-title":"Asian Conference on Computer Vision. Springer, 235--250","author":"Liu Yi-Chieh","year":"2018","unstructured":"Yi-Chieh Liu , Hao-Hsiang Yang , C-H Huck Yang , Jia-Hong Huang , Meng Tian , Hiromasa Morikawa , Yi-Chang James Tsai , and Jesper Tegner . 2018 . Synthesizing new retinal symptom images by multiple generative models . In Asian Conference on Computer Vision. Springer, 235--250 . Yi-Chieh Liu, Hao-Hsiang Yang, C-H Huck Yang, Jia-Hong Huang, Meng Tian, Hiromasa Morikawa, Yi-Chang James Tsai, and Jesper Tegner. 2018. Synthesizing new retinal symptom images by multiple generative models. In Asian Conference on Computer Vision. Springer, 235--250."},{"key":"e_1_3_2_1_40_1","volume-title":"A multi-world approach to question answering about real-world scenes based on uncertain input. arXiv preprint arXiv:1410.0210","author":"Malinowski Mateusz","year":"2014","unstructured":"Mateusz Malinowski and Mario Fritz . 2014. A multi-world approach to question answering about real-world scenes based on uncertain input. arXiv preprint arXiv:1410.0210 ( 2014 ). Mateusz Malinowski and Mario Fritz. 2014. A multi-world approach to question answering about real-world scenes based on uncertain input. arXiv preprint arXiv:1410.0210 (2014)."},{"key":"e_1_3_2_1_41_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2015.9"},{"key":"e_1_3_2_1_42_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-017-1038-2"},{"key":"e_1_3_2_1_43_1","volume-title":"Distributed representations of words and phrases and their compositionality. arXiv preprint arXiv:1310.4546","author":"Mikolov Tomas","year":"2013","unstructured":"Tomas Mikolov , Ilya Sutskever , Kai Chen , Greg Corrado , and Jeffrey Dean . 2013. Distributed representations of words and phrases and their compositionality. arXiv preprint arXiv:1310.4546 ( 2013 ). Tomas Mikolov, Ilya Sutskever, Kai Chen, Greg Corrado, and Jeffrey Dean. 2013. Distributed representations of words and phrases and their compositionality. arXiv preprint arXiv:1310.4546 (2013)."},{"key":"e_1_3_2_1_44_1","unstructured":"M Niemeijer X Xu A Dumitrescu P Gupta B van Ginneken J Folk and M Abramoff. 2011. INSPIRE-AVR: Iowa normative set for processing images of the retina-artery vein ratio.  M Niemeijer X Xu A Dumitrescu P Gupta B van Ginneken J Folk and M Abramoff. 2011. INSPIRE-AVR: Iowa normative set for processing images of the retina-artery vein ratio."},{"key":"e_1_3_2_1_45_1","doi-asserted-by":"publisher","DOI":"10.1155\/2009\/235746"},{"key":"e_1_3_2_1_46_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.01098"},{"key":"e_1_3_2_1_47_1","volume-title":"Proceedings of ACL. Association for Computational Linguistics, 311--318","author":"Papineni Kishore","year":"2002","unstructured":"Kishore Papineni , Salim Roukos , Todd Ward , and Wei-Jing Zhu . 2002 . BLEU: a method for automatic evaluation of machine translation . In Proceedings of ACL. Association for Computational Linguistics, 311--318 . Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002. BLEU: a method for automatic evaluation of machine translation. In Proceedings of ACL. Association for Computational Linguistics, 311--318."},{"key":"e_1_3_2_1_48_1","doi-asserted-by":"publisher","DOI":"10.3115\/v1\/D14-1162"},{"key":"e_1_3_2_1_49_1","volume-title":"Deep contextualized word representations. arXiv preprint arXiv:1802.05365","author":"Peters Matthew E","year":"2018","unstructured":"Matthew E Peters , Mark Neumann , Mohit Iyyer , Matt Gardner , Christopher Clark , Kenton Lee , and Luke Zettlemoyer . 2018. Deep contextualized word representations. arXiv preprint arXiv:1802.05365 ( 2018 ). Matthew E Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer. 2018. Deep contextualized word representations. arXiv preprint arXiv:1802.05365 (2018)."},{"key":"e_1_3_2_1_50_1","doi-asserted-by":"publisher","DOI":"10.3390\/data3030025"},{"key":"e_1_3_2_1_51_1","volume-title":"Language models are unsupervised multitask learners. OpenAI blog","author":"Radford Alec","year":"2019","unstructured":"Alec Radford , Jeffrey Wu , Rewon Child , David Luan , Dario Amodei , and Ilya Sutskever . 2019. Language models are unsupervised multitask learners. OpenAI blog , Vol. 1 , 8 ( 2019 ), 9. Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019. Language models are unsupervised multitask learners. OpenAI blog , Vol. 1, 8 (2019), 9."},{"key":"e_1_3_2_1_52_1","volume-title":"et almbox","author":"Russakovsky Olga","year":"2015","unstructured":"Olga Russakovsky , Jia Deng , Hao Su , Jonathan Krause , Sanjeev Satheesh , Sean Ma , Zhiheng Huang , Andrej Karpathy , Aditya Khosla , Michael Bernstein , et almbox . 2015 . Imagenet large scale visual recognition challenge. International journal of computer vision , Vol. 115 , 3 (2015), 211--252. Olga Russakovsky, Jia Deng, Hao Su, Jonathan Krause, Sanjeev Satheesh, Sean Ma, Zhiheng Huang, Andrej Karpathy, Aditya Khosla, Michael Bernstein, et almbox. 2015. Imagenet large scale visual recognition challenge. International journal of computer vision , Vol. 115, 3 (2015), 211--252."},{"key":"e_1_3_2_1_53_1","unstructured":"Sam Scott and Stan Matwin. 1998. Text classification using WordNet hypernyms. In Usage of WordNet in Natural Language Processing Systems .  Sam Scott and Stan Matwin. 1998. Text classification using WordNet hypernyms. In Usage of WordNet in Natural Language Processing Systems ."},{"key":"e_1_3_2_1_54_1","volume-title":"Very deep convolutional networks for large-scale image recognition. arXiv preprint arXiv:1409.1556","author":"Simonyan Karen","year":"2014","unstructured":"Karen Simonyan and Andrew Zisserman . 2014. Very deep convolutional networks for large-scale image recognition. arXiv preprint arXiv:1409.1556 ( 2014 ). Karen Simonyan and Andrew Zisserman. 2014. Very deep convolutional networks for large-scale image recognition. arXiv preprint arXiv:1409.1556 (2014)."},{"key":"e_1_3_2_1_55_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISBI.2014.6867807"},{"key":"e_1_3_2_1_56_1","doi-asserted-by":"publisher","DOI":"10.9790\/0661-16153438"},{"key":"e_1_3_2_1_57_1","doi-asserted-by":"publisher","DOI":"10.1109\/TMI.2004.825627"},{"key":"e_1_3_2_1_58_1","volume-title":"Attention is all you need. arXiv preprint arXiv:1706.03762","author":"Vaswani Ashish","year":"2017","unstructured":"Ashish Vaswani , Noam Shazeer , Niki Parmar , Jakob Uszkoreit , Llion Jones , Aidan N Gomez , Lukasz Kaiser , and Illia Polosukhin . 2017. Attention is all you need. arXiv preprint arXiv:1706.03762 ( 2017 ). Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Lukasz Kaiser, and Illia Polosukhin. 2017. Attention is all you need. arXiv preprint arXiv:1706.03762 (2017)."},{"key":"e_1_3_2_1_59_1","volume-title":"Maria Ant\u00f2nia Barcel\u00f3, and Marc Saez","author":"V\u00e1zquez SG","year":"2013","unstructured":"SG V\u00e1zquez , Brais Cancela , Noelia Barreira , Manuel G Penedo , M Rodr'iguez-Blanco , M Pena Seijo , G Coll de Tuero , Maria Ant\u00f2nia Barcel\u00f3, and Marc Saez . 2013 . Improving retinal artery and vein classification by means of a minimal path approach. Machine vision and applications , Vol. 24 , 5 (2013), 919--930. SG V\u00e1zquez, Brais Cancela, Noelia Barreira, Manuel G Penedo, M Rodr'iguez-Blanco, M Pena Seijo, G Coll de Tuero, Maria Ant\u00f2nia Barcel\u00f3, and Marc Saez. 2013. Improving retinal artery and vein classification by means of a minimal path approach. Machine vision and applications , Vol. 24, 5 (2013), 919--930."},{"key":"e_1_3_2_1_60_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2015.7299087"},{"key":"e_1_3_2_1_61_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2015.7298935"},{"key":"e_1_3_2_1_62_1","volume-title":"ICML Workshop on Computational Biology","author":"Huck Yang C-H","year":"2018","unstructured":"C-H Huck Yang , Jia-Hong Huang , Fangyu Liu , Fang-Yi Chiu , Mengya Gao , Weifeng Lyu , Jesper Tegner , 2018 a. A novel hybrid machine learning model for auto-classification of retinal diseases . ICML Workshop on Computational Biology (2018). C-H Huck Yang, Jia-Hong Huang, Fangyu Liu, Fang-Yi Chiu, Mengya Gao, Weifeng Lyu, Jesper Tegner, et almbox. 2018a. A novel hybrid machine learning model for auto-classification of retinal diseases. ICML Workshop on Computational Biology (2018)."},{"key":"e_1_3_2_1_63_1","volume-title":"Asian Conference on Computer Vision. Springer, 323--338","author":"Huck Yang C-H","year":"2018","unstructured":"C-H Huck Yang , Fangyu Liu , Jia-Hong Huang , Meng Tian , MD I- Hung Lin , Yi Chieh Liu , Hiromasa Morikawa , Hao-Hsiang Yang , and Jesper Tegner . 2018 b. Auto-classification of retinal diseases in the limit of sparse data using a two-streams machine learning model . In Asian Conference on Computer Vision. Springer, 323--338 . C-H Huck Yang, Fangyu Liu, Jia-Hong Huang, Meng Tian, MD I-Hung Lin, Yi Chieh Liu, Hiromasa Morikawa, Hao-Hsiang Yang, and Jesper Tegner. 2018b. Auto-classification of retinal diseases in the limit of sparse data using a two-streams machine learning model. In Asian Conference on Computer Vision. Springer, 323--338."}],"event":{"name":"ICMR '21: International Conference on Multimedia Retrieval","location":"Taipei Taiwan","acronym":"ICMR '21","sponsor":["SIGMM ACM Special Interest Group on Multimedia"]},"container-title":["Proceedings of the 2021 International Conference on Multimedia Retrieval"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3460426.3463667","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3460426.3463667","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T20:17:04Z","timestamp":1750191424000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3460426.3463667"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2021,8,24]]},"references-count":63,"alternative-id":["10.1145\/3460426.3463667","10.1145\/3460426"],"URL":"https:\/\/doi.org\/10.1145\/3460426.3463667","relation":{},"subject":[],"published":{"date-parts":[[2021,8,24]]},"assertion":[{"value":"2021-09-01","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}