{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,28]],"date-time":"2026-07-28T02:59:40Z","timestamp":1785207580889,"version":"3.55.0"},"reference-count":177,"publisher":"Association for Computing Machinery (ACM)","issue":"2","funder":[{"DOI":"10.13039\/501100001807","name":"Funda\u00e7\u00e3o de Amparo \u00e0 Pesquisa do Estado de S\u00e3o Paulo","doi-asserted-by":"crossref","award":["21\/14725-3 and 23\/12493-3"],"award-info":[{"award-number":["21\/14725-3 and 23\/12493-3"]}],"id":[{"id":"10.13039\/501100001807","id-type":"DOI","asserted-by":"crossref"}]},{"name":"Conselho Nacional de Desenvolvimento Cient\u00edfico e Tecnol\u00f3gico (CNPQ), Swiss National Science Foundation","award":["200021E_214653"],"award-info":[{"award-number":["200021E_214653"]}]},{"name":"Santos Dumont supercomputer at the LNCC"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Comput. Healthcare"],"published-print":{"date-parts":[[2026,4,30]]},"abstract":"<jats:p>Ensuring equitable Artificial Intelligence (AI) in healthcare demands systems that make unbiased decisions across all demographic groups, bridging technical innovation with ethical principles. Foundation Models (FMs), trained on vast datasets through self-supervised learning, enable efficient adaptation across medical imaging tasks while reducing dependency on labeled data. These models demonstrate potential for enhancing fairness, though significant challenges remain in achieving consistent performance across demographic groups. Our review indicates that effective bias mitigation in FMs requires systematic interventions throughout all stages of development. While previous approaches focused primarily on model-level bias mitigation, our analysis reveals that fairness in FMs requires integrated interventions throughout the development pipeline, from data documentation to deployment protocols. This comprehensive framework advances current knowledge by demonstrating how systematic bias mitigation, combined with policy engagement, can effectively address both technical and institutional barriers to equitable AI in healthcare. The development of equitable FMs represents a critical step toward democratizing advanced healthcare technologies, particularly for underserved populations and regions with limited medical infrastructure and computational resources.<\/jats:p>","DOI":"10.1145\/3793542","type":"journal-article","created":{"date-parts":[[2026,1,24]],"date-time":"2026-01-24T11:25:28Z","timestamp":1769253928000},"page":"1-24","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":3,"title":["Fair Foundation Models for Medical Image Analysis: Challenges and Perspectives"],"prefix":"10.1145","volume":"7","author":[{"ORCID":"https:\/\/orcid.org\/0009-0005-9526-9583","authenticated-orcid":false,"given":"Dilermando","family":"Queiroz","sequence":"first","affiliation":[{"name":"Federal University of S\u00e3o Paulo, S\u00e3o Jos\u00e9 dos Campos, Brazil"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0001-4637-0301","authenticated-orcid":false,"given":"Anderson","family":"Carlos","sequence":"additional","affiliation":[{"name":"Federal Institute of Goi\u00e1s, Goi\u00e2nia, Brazil"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-7248-4014","authenticated-orcid":false,"given":"Andr\u00e9","family":"Anjos","sequence":"additional","affiliation":[{"name":"Idiap Research Institute, Martigny, Switzerland"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-1397-6005","authenticated-orcid":false,"given":"Lilian","family":"Berton","sequence":"additional","affiliation":[{"name":"Federal University of S\u00e3o Paulo, S\u00e3o Jos\u00e9 dos Campos, Brazil"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2026,3,17]]},"reference":[{"key":"e_1_3_2_2_2","unstructured":"TOP500. [n.d.]. November 2024. Retrieved from https:\/\/top500.org\/lists\/top500\/2024\/11\/"},{"key":"e_1_3_2_3_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41746-024-01232-3"},{"key":"e_1_3_2_4_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-08215-3_5"},{"key":"e_1_3_2_5_2","unstructured":"Ibrahim Alabdulmohsin Xiao Wang Andreas Steiner Priya Goyal Alexander D\u2019Amour and Xiaohua Zhai. 2024. CLIP the bias: How useful is balancing data in multimodal learning? arXiv:2403.04547. Retrieved from https:\/\/arxiv.org\/abs\/2403.04547"},{"key":"e_1_3_2_6_2","doi-asserted-by":"publisher","DOI":"10.1109\/MIC.2022.3147923"},{"key":"e_1_3_2_7_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41598-024-79863-5"},{"key":"e_1_3_2_8_2","doi-asserted-by":"crossref","unstructured":"McKane Andrus and Sarah Villeneuve. 2022. Demographic-reliant algorithmic fairness: Characterizing the risks of demographic data collection in the pursuit of fairness. arXiv:2205.01038. Retrieved from https:\/\/arxiv.org\/abs\/2205.01038","DOI":"10.1145\/3531146.3533226"},{"key":"e_1_3_2_9_2","doi-asserted-by":"publisher","DOI":"10.1145\/3617694.3623234"},{"key":"e_1_3_2_10_2","doi-asserted-by":"crossref","unstructured":"Mahmoud Assran Quentin Duval Ishan Misra Piotr Bojanowski Pascal Vincent Michael Rabbat Yann LeCun and Nicolas Ballas. 2023. Self-supervised learning from images with a joint-embedding predictive architecture. arXiv:2301.08243. Retrieved from https:\/\/arxiv.org\/abs\/2301.08243","DOI":"10.1109\/CVPR52729.2023.01499"},{"key":"e_1_3_2_11_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41551-023-01049-7"},{"key":"e_1_3_2_12_2","unstructured":"Eugene Bagdasaryan and Vitaly Shmatikov. 2019. Differential privacy has disparate impact on model accuracy. arXiv:1905.12101. Retrieved from https:\/\/arxiv.org\/abs\/1905.12101"},{"key":"e_1_3_2_13_2","doi-asserted-by":"publisher","DOI":"10.1016\/S0140-6736(17)30569-X"},{"key":"e_1_3_2_14_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2025.acl-long.914"},{"key":"e_1_3_2_15_2","volume-title":"Principles of Biomedical Ethics","author":"Beauchamp Tom L.","year":"1994","unstructured":"Tom L. Beauchamp and James F. Childress. 1994. Principles of Biomedical Ethics. Edicoes Loyola."},{"key":"e_1_3_2_16_2","unstructured":"Y. Bengio S\u00f6ren Mindermann Daniel Privitera Tamay Besiroglu Rishi Bommasani Stephen Casper Yejin Choi Philip Fox Ben Garfinkel Danielle Goldfarb et al. 2025. International AI safety report. arXiv:2501.17805. Retrieved from https:\/\/arxiv.org\/abs\/2501.17805"},{"key":"e_1_3_2_17_2","doi-asserted-by":"publisher","DOI":"10.2147\/PPA.S87928"},{"key":"e_1_3_2_18_2","unstructured":"Rishi Bommasani Drew A. Hudson Ehsan Adeli Russ Altman Simran Arora Sydney von Arx Michael S. Bernstein Jeannette Bohg Antoine Bosselut Emma Brunskill et al. 2022. On the opportunities and risks of foundation models. arXiv:2108.07258. Retrieved from https:\/\/arxiv.org\/abs\/2108.07258"},{"key":"e_1_3_2_19_2","doi-asserted-by":"publisher","DOI":"10.1126\/science.adp1848"},{"key":"e_1_3_2_20_2","unstructured":"Rishi Bommasani Kevin Klyman Shayne Longpre Sayash Kapoor Nestor Maslej Betty Xiong Daniel Zhang and Percy Liang. 2023. The foundation model transparency index. arXiv:2310.12941. Retrieved from https:\/\/arxiv.org\/abs\/2310.12941"},{"key":"e_1_3_2_21_2","doi-asserted-by":"publisher","DOI":"10.1145\/3462244.3479897"},{"key":"e_1_3_2_22_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41597-020-00622-y"},{"key":"e_1_3_2_23_2","doi-asserted-by":"publisher","DOI":"10.1167\/tvst.10.2.13"},{"key":"e_1_3_2_24_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.media.2020.101797"},{"key":"e_1_3_2_25_2","first-page":"3992","article-title":"Optimized pre-processing for discrimination prevention","volume":"30","author":"Calmon Flavio","year":"2017","unstructured":"Flavio Calmon, Dennis Wei, Bhanukiran Vinzamuri, Karthikeyan Natesan Ramamurthy, and Kush R. Varshney. 2017. Optimized pre-processing for discrimination prevention. In Advances in Neural Information Processing Systems, Vol. 30, 3992\u20134001.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_26_2","doi-asserted-by":"publisher","DOI":"10.1515\/bmt-2023-0325"},{"key":"e_1_3_2_27_2","unstructured":"L. Elisa Celis and Vijay Keswani. 2019. Improved adversarial learning for fair classification. arXiv:1901.10443. Retrieved from https:\/\/arxiv.org\/abs\/1901.10443"},{"key":"e_1_3_2_28_2","unstructured":"Delong Chen Samuel Cahyawijaya Etsuko Ishii Ho Shu Chan Yejin Bang and Pascale Fung. 2024. What makes for good image captions? arXiv:2405.00485. Retrieved from https:\/\/arxiv.org\/abs\/2405.00485"},{"key":"e_1_3_2_29_2","unstructured":"Delong Chen Theo Moutakanni Willy Chung Yejin Bang Ziwei Ji Allen Bolourchi and Pascale Fung. 2025. Planning with reasoning using vision language world model. arXiv:2509.02722. Retrieved from https:\/\/arxiv.org\/abs\/2509.02722"},{"key":"e_1_3_2_30_2","doi-asserted-by":"publisher","DOI":"10.1146\/annurev-biodatasci-092820-114757"},{"key":"e_1_3_2_31_2","unstructured":"Jiawei Chen Dingkang Yang Tong Wu Yue Jiang Xiaolu Hou Mingcheng Li Shunli Wang Dongling Xiao Ke Li and Lihua Zhang. 2024. Detecting and evaluating medical hallucinations in large vision language models. arXiv:2406.10185. Retrieved from https:\/\/arxiv.org\/abs\/2406.10185"},{"key":"e_1_3_2_32_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41551-023-01056-8"},{"key":"e_1_3_2_33_2","unstructured":"Shan Chen Jack Gallifant Mingye Gao Pedro Moreira Nikolaj Munch Ajay Muthukkumar Arvind Rajan Jaya Kolluri Amelia Fiske Janna Hastings et al. 2024. Cross-care: Assessing the healthcare implications of pre-training data on language model bias. arXiv:2405.05506. Retrieved from https:\/\/arxiv.org\/abs\/2405.05506"},{"key":"e_1_3_2_34_2","unstructured":"Ting Chen Simon Kornblith Mohammad Norouzi and Geoffrey Hinton. 2020. A simple framework for contrastive learning of visual representations. arXiv:2002.05709. Retrieved from https:\/\/arxiv.org\/abs\/2002.05709"},{"key":"e_1_3_2_35_2","volume-title":"Proceedings of the 1st Conference on Fairness, Accountability and Transparency","author":"Chouldechova Alexandra","year":"2018","unstructured":"Alexandra Chouldechova, Diana Benavides-Prado, Oleksandr Fialko, and Rhema Vaithianathan. 2018. A case study of algorithm-assisted decision making in child maltreatment hotline screening decisions. In Proceedings of the 1st Conference on Fairness, Accountability and Transparency. PMLR."},{"key":"e_1_3_2_36_2","doi-asserted-by":"publisher","DOI":"10.1136\/bmj-2023-078378"},{"key":"e_1_3_2_37_2","unstructured":"Ben Cottier Robi Rahman Loredana Fattorini Nestor Maslej and David Owen. 2024. The rising costs of training frontier AI models. arXiv:2405.21015. Retrieved from https:\/\/arxiv.org\/abs\/2405.21015"},{"key":"e_1_3_2_38_2","unstructured":"Jiequan Cui Beier Zhu Xin Wen Xiaojuan Qi Bei Yu and Hanwang Zhang. 2024. Classes are not equal: An empirical study on image recognition fairness. arXiv:2402.18133. Retrieved from https:\/\/arxiv.org\/abs\/2402.18133"},{"key":"e_1_3_2_39_2","unstructured":"Maria Correia de Verdier Rachit Saluja Louis Gagnon Dominic LaBella Ujjwall Baid Nourel Hoda Tahon Martha Foltyn-Dumitru Jikai Zhang Maram Alafif Saif Baig et al. 2024. The 2024 brain tumor segmentation (BraTS) challenge: Glioma segmentation on post-treatment MRI. arXiv:2405.18368. Retrieved from https:\/\/arxiv.org\/abs\/2405.18368"},{"key":"e_1_3_2_40_2","doi-asserted-by":"crossref","unstructured":"Sepehr Dehdashtian Bashir Sadeghi and Vishnu Naresh Boddeti. 2024. Utility-fairness trade-offs and how to find them. arXiv:2404.09454. Retrieved from https:\/\/arxiv.org\/abs\/2404.09454","DOI":"10.1109\/CVPR52733.2024.01144"},{"key":"e_1_3_2_41_2","unstructured":"Matt Deitke Christopher Clark Sangho Lee Rohun Tripathi Yue Yang Jae Sung Park Mohammadreza Salehi Niklas Muennighoff Kyle Lo Luca Soldaini et al. 2024. Molmo and PixMo: Open weights and open data for state-of-the-art multimodal models. arXiv:2409.17146. Retrieved from https:\/\/arxiv.org\/abs\/2409.17146"},{"key":"e_1_3_2_42_2","unstructured":"Mohammad Mahdi Derakhshani Dheeraj Varghese Marzieh Fadaee and Cees G. M. Snoek. 2025. NeoBabel: A multilingual open tower for visual generation. arXiv:2507.06137. Retrieved from https:\/\/arxiv.org\/abs\/2507.06137"},{"key":"e_1_3_2_43_2","doi-asserted-by":"crossref","unstructured":"Tim Dettmers Artidoro Pagnoni Ari Holtzman and Luke Zettlemoyer. 2023. QLoRA: Efficient finetuning of quantized LLMs. arXiv:2305.14314. Retrieved from https:\/\/arxiv.org\/abs\/2305.14314","DOI":"10.52202\/075280-0441"},{"key":"e_1_3_2_44_2","unstructured":"Zhoujie Ding Ken Ziyu Liu Pura Peetathawatchai Berivan Isik and Sanmi Koyejo. 2024. On fairness of low-rank adaptation of large models. arXiv:2405. Retrieved from https:\/\/arxiv.org\/abs\/2405.17512"},{"key":"e_1_3_2_45_2","unstructured":"Alexey Dosovitskiy Lucas Beyer Alexander Kolesnikov Dirk Weissenborn Xiaohua Zhai Thomas Unterthiner Mostafa Dehghani Matthias Minderer Georg Heigold Sylvain Gelly et al. 2021. An image is worth 16x16 words: Transformers for image recognition at scale. arXiv:2010.11929. Retrieved from https:\/\/arxiv.org\/abs\/2010.11929"},{"key":"e_1_3_2_46_2","doi-asserted-by":"publisher","DOI":"10.1109\/MIS.2020.3000681"},{"key":"e_1_3_2_47_2","unstructured":"Raman Dutt Ondrej Bohdal Sotirios A. Tsaftaris and Timothy Hospedales. 2024. FairTune: Optimizing parameter efficient fine tuning for fairness in medical image analysis. arXiv:2310.05055. Retrieved from https:\/\/arxiv.org\/abs\/2310.05055"},{"key":"e_1_3_2_48_2","unstructured":"European Parliament and Council. 2024. Regulation (EU) 2024\/1689 of the European Parliament and of the Council of 13 June 2024 laying down harmonised rules on artificial intelligence (Artificial Intelligence Act) OJ L 2024\/1689. Retrieved July 12 2024 from http:\/\/data.europa.eu\/eli\/reg\/2024\/1689\/oj"},{"key":"e_1_3_2_49_2","doi-asserted-by":"crossref","unstructured":"Talfan Evans Nikhil Parthasarathy Hamza Merzic and Olivier J. Henaff. 2024. Data curation via joint example selection further accelerates multimodal learning. arXiv:2406.17711. Retrieved from https:\/\/arxiv.org\/abs\/2406.17711","DOI":"10.52202\/079017-4485"},{"key":"e_1_3_2_50_2","doi-asserted-by":"crossref","unstructured":"Talfan Evans Shreya Pathak Hamza Merzic Jonathan Schwarz Ryutaro Tanno and Olivier J. Henaff. 2024. Bad students make great teachers: Active learning accelerates large-scale visual understanding. arXiv:2312.05328. Retrieved from https:\/\/arxiv.org\/abs\/2312.05328","DOI":"10.1007\/978-3-031-72643-9_16"},{"key":"e_1_3_2_51_2","unstructured":"Health Data Research Innovation Gateway. 2024. Moorfields Eye Image BioResource 001. Health Data Research Innovation Gateway."},{"key":"e_1_3_2_52_2","doi-asserted-by":"publisher","DOI":"10.1038\/s42256-020-00257-z"},{"key":"e_1_3_2_53_2","unstructured":"Zhengyang Geng Mingyang Deng Xingjian Bai J. Zico Kolter and Kaiming He. 2025. Mean flows for one-step generative modeling. arXiv:2505.13447. Retrieved from https:\/\/arxiv.org\/abs\/2505.13447"},{"key":"e_1_3_2_54_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.ebiom.2023.104467"},{"key":"e_1_3_2_55_2","doi-asserted-by":"publisher","DOI":"10.1148\/ryai.230060"},{"key":"e_1_3_2_56_2","unstructured":"Priya Goyal Mathilde Caron Benjamin Lefaudeux Min Xu Pengchao Wang Vivek Pai Mannat Singh Vitaliy Liptchinsky Ishan Misra Armand Joulin and Piotr Bojanowski. 2021. Self-supervised pretraining of visual features in the wild. arXiv:2103.01988. Retrieved from https:\/\/arxiv.org\/abs\/2103.01988"},{"key":"e_1_3_2_57_2","unstructured":"Priya Goyal Quentin Duval Isaac Seessel Mathilde Caron Ishan Misra Levent Sagun Armand Joulin and Piotr Bojanowski. 2022. Vision models are more robust and fair when pretrained on uncurated images without supervision. arXiv:2202. Retrieved from https:\/\/arxiv.org\/abs\/2202.08360"},{"key":"e_1_3_2_58_2","doi-asserted-by":"crossref","unstructured":"Matthew Groh Caleb Harris Luis Soenksen Felix Lau Rachel Han Aerin Kim Arash Koochek and Omar Badri. 2021. Evaluating deep neural networks trained on clinical images in dermatology with the fitzpatrick 17k dataset. arXiv:2104.09957. Retrieved from https:\/\/arxiv.org\/abs\/2104.09957","DOI":"10.1109\/CVPRW53098.2021.00201"},{"key":"e_1_3_2_59_2","unstructured":"Zishan Gu Changchang Yin Fenglin Liu and Ping Zhang. 2024. MedVH: Towards systematic evaluation of hallucination for large vision language models in the medical context. arXiv:2407.02730. Retrieved from https:\/\/arxiv.org\/abs\/2407.02730"},{"key":"e_1_3_2_60_2","doi-asserted-by":"publisher","DOI":"10.1001\/jama.2016.17216"},{"key":"e_1_3_2_61_2","doi-asserted-by":"crossref","unstructured":"Melissa Hall Laura Gustafson Aaron Adcock Ishan Misra and Candace Ross. 2023. Vision-language models performing zero-shot tasks exhibit gender-based disparities. arXiv:2301.11100. Retrieved from https:\/\/arxiv.org\/abs\/2301.11100","DOI":"10.1109\/ICCVW60793.2023.00294"},{"key":"e_1_3_2_62_2","unstructured":"Kaiming He Xinlei Chen Saining Xie Yanghao Li Piotr Doll\u00e1r and Ross Girshick. 2021. Masked autoencoders are scalable vision learners. arXiv:2111.06377. Retrieved from https:\/\/arxiv.org\/abs\/2111.06377"},{"key":"e_1_3_2_63_2","unstructured":"Stefan Hegselmann Shannon Zejiang Shen Florian Gierse Monica Agrawal David Sontag and Xiaoyi Jiang. 2024. A data-centric approach to generate faithful and high quality patient summaries with large language models. arXiv:2402.15422. Retrieved from https:\/\/arxiv.org\/abs\/2402.15422"},{"key":"e_1_3_2_64_2","doi-asserted-by":"crossref","unstructured":"Greg Heinrich Mike Ranzinger Hongxu Yin Yao Lu Jan Kautz Andrew Tao Bryan Catanzaro and Pavlo Molchanov. 2025. RADIOv2.5: Improved baselines for agglomerative vision foundation models. arXiv:2412.07679. Retrieved from https:\/\/arxiv.org\/abs\/2412.07679","DOI":"10.1109\/CVPR52734.2025.02094"},{"key":"e_1_3_2_65_2","volume-title":"Ethics Guidelines for Trustworthy AI","author":"High-Level Expert Group on Artificial Intelligence","year":"2019","unstructured":"High-Level Expert Group on Artificial Intelligence. 2019. Ethics Guidelines for Trustworthy AI. Retrieved from https:\/\/digital-strategy.ec.europa.eu\/en\/library\/ethics-guidelines-trustworthy-ai"},{"key":"e_1_3_2_66_2","unstructured":"Edward J. Hu Yelong Shen Phillip Wallis Zeyuan Allen-Zhu Yuanzhi Li Shean Wang Lu Wang and Weizhu Chen. 2021. LoRA: Low-rank adaptation of large language models. arXiv:2106.09685. Retrieved from https:\/\/arxiv.org\/abs\/2106.09685"},{"key":"e_1_3_2_67_2","doi-asserted-by":"publisher","DOI":"10.1145\/3703155"},{"key":"e_1_3_2_68_2","unstructured":"Ben Hutchinson Jason Baldridge and Vinodkumar Prabhakaran. 2022. Underspecification in scene description-to-depiction tasks. arXiv:2210.05815. Retrieved from https:\/\/arxiv.org\/abs\/2210.05815"},{"key":"e_1_3_2_69_2","doi-asserted-by":"crossref","unstructured":"Jeremy Irvin Pranav Rajpurkar Michael Ko Yifan Yu Silviana Ciurea-Ilcus Chris Chute Henrik Marklund Behzad Haghgoo Robyn Ball Katie Shpanskaya et al. 2019. CheXpert: A large chest radiograph dataset with uncertainty labels and expert comparison. arXiv:1901.07031. Retrieved from https:\/\/arxiv.org\/abs\/1901.07031","DOI":"10.1609\/aaai.v33i01.3301590"},{"key":"e_1_3_2_70_2","doi-asserted-by":"publisher","DOI":"10.1007\/s11633-022-1371-y"},{"key":"e_1_3_2_71_2","doi-asserted-by":"publisher","DOI":"10.1145\/3571730"},{"key":"e_1_3_2_72_2","unstructured":"Ruinan Jin Wenlong Deng Minghui Chen and Xiaoxiao Li. 2024. Universal debiased editing on foundation models for fair medical image classification. arXiv:2403.06104. Retrieved from https:\/\/arxiv.org\/abs\/2403.06104"},{"key":"e_1_3_2_73_2","first-page":"111318","article-title":"FairMedFM: Fairness benchmarking for medical imaging foundation models","volume":"37","author":"Jin Ruinan","year":"2024","unstructured":"Ruinan Jin, Zikang Xu, Yuan Zhong, Qingsong Yao, Qi Dou, S. K. Zhou, and Xiaoxiao Li. 2024. FairMedFM: Fairness benchmarking for medical imaging foundation models. In Advances in Neural Information Processing Systems, Vol. 37, 111318\u2013111357","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_74_2","doi-asserted-by":"crossref","unstructured":"Alistair E. W. Johnson Tom J. Pollard Nathaniel R. Greenbaum Matthew P. Lungren Chih-Ying Deng Yifan Peng Zhiyong Lu Roger G. Mark Seth J. Berkowitz and Steven Horng. 2019. MIMIC-CXR-JPG a large publicly available database of labeled chest radiographs. arXiv:1901.07042. Retrieved from https:\/\/arxiv.org\/abs\/1901.07042","DOI":"10.1038\/s41597-019-0322-0"},{"key":"e_1_3_2_75_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-33486-33"},{"key":"e_1_3_2_76_2","doi-asserted-by":"publisher","DOI":"10.1038\/s42256-023-00652-2"},{"key":"e_1_3_2_77_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-44917-8_25"},{"key":"e_1_3_2_78_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.cell.2018.02.010"},{"key":"e_1_3_2_79_2","volume-title":"Proceedings of the 3rd Machine Learning for Health Symposium","author":"Khan Muhammad Osama","year":"2023","unstructured":"Muhammad Osama Khan, Muhammad Muneeb Afzal, Shujaat Mirza, and Yi Fang. 2023. How fair are medical imaging foundation models? In Proceedings of the 3rd Machine Learning for Health Symposium. PMLR."},{"key":"e_1_3_2_80_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.ebiom.2024.105174"},{"key":"e_1_3_2_81_2","doi-asserted-by":"crossref","unstructured":"Yubin Kim Hyewon Jeong Shan Chen Shuyue Stella Li Chanwoo Park Mingyu Lu Kumail Alhamoud Jimin Mun Cristina Grau Minseok Jung et al. 2025. Medical hallucinations in foundation models and their impact on healthcare. arXiv:2503.05777. Retrieved from https:\/\/arxiv.org\/abs\/2503.05777","DOI":"10.1101\/2025.02.28.25323115"},{"key":"e_1_3_2_82_2","doi-asserted-by":"crossref","unstructured":"Alexander Kirillov Eric Mintun Nikhila Ravi Hanzi Mao Chloe Rolland Laura Gustafson Tete Xiao Spencer Whitehead Alexander C. Berg Wan-Yen Lo et al. 2023. Segment anything. arXiv:2304.02643. Retrieved from https:\/\/arxiv.org\/abs\/2304.02643","DOI":"10.1109\/ICCV51070.2023.00371"},{"key":"e_1_3_2_83_2","unstructured":"Ira Ktena Olivia Wiles Isabela Albuquerque Sylvestre-Alvise Rebuffi Ryutaro Tanno Abhijit Guha Roy Shekoofeh Azizi Danielle Belgrave Pushmeet Kohli Alan Karthikesalingam et al. 2023. Generative models improve fairness of medical classifiers under distribution shifts. arXiv:2304.09218. Retrieved from https:\/\/arxiv.org\/abs\/2304.09218"},{"key":"e_1_3_2_84_2","unstructured":"Andrew Kyle Lampinen Stephanie C. Y. Chan and Katherine Hermann. 2024. Learned feature representations are biased by complexity learning order position and more. arXiv:2405.05847. Retrieved from https:\/\/arxiv.org\/abs\/2405.05847"},{"key":"e_1_3_2_85_2","unstructured":"Yann LeCun. 2022. A Path Towards Autonomous Machine Intelligence. Retrieved from https:\/\/openreview.net\/pdf?id=BZ5a1r-kVsf"},{"key":"e_1_3_2_86_2","unstructured":"Katherine Lee Daphne Ippolito Andrew Nystrom Chiyuan Zhang Douglas Eck Chris Callison-Burch and Nicholas Carlini. 2022. Deduplicating training data makes language models better. arXiv:2107.06499. Retrieved from https:\/\/arxiv.org\/abs\/2107.06499"},{"key":"e_1_3_2_87_2","doi-asserted-by":"publisher","DOI":"10.1136\/bmj-2024-081554"},{"key":"e_1_3_2_88_2","unstructured":"Chunyuan Li Cliff Wong Sheng Zhang Naoto Usuyama Haotian Liu Jianwei Yang Tristan Naumann Hoifung Poon and Jianfeng Gao. 2023. LLaVA-Med: Training a large language-and-vision assistant for biomedicine in one day. arXiv:2306.00890. Retrieved from https:\/\/arxiv.org\/abs\/2306.00890"},{"key":"e_1_3_2_89_2","doi-asserted-by":"publisher","DOI":"10.1093\/bib\/bbae548"},{"key":"e_1_3_2_90_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-72390-2_41"},{"key":"e_1_3_2_91_2","unstructured":"Weixiong Lin Ziheng Zhao Xiaoman Zhang Chaoyi Wu Ya Zhang Yanfeng Wang and Weidi Xie. 2023. PMC-CLIP: Contrastive language-image pre-training using biomedical documents. arXiv:2303.07240. Retrieved from https:\/\/arxiv.org\/abs\/2303.07240"},{"key":"e_1_3_2_92_2","volume-title":"Proceedings of the 38th Annual Conference on Neural Information Processing Systems","author":"Littwin Etai","year":"2024","unstructured":"Etai Littwin, Omid Saremi, Madhu Advani, Vimal Thilak, Preetum Nakkiran, Chen Huang, and Joshua M. Susskind. 2024. How JEPA avoids noisy features: The implicit bias of deep linear self distillation networks. In Proceedings of the 38th Annual Conference on Neural Information Processing Systems."},{"key":"e_1_3_2_93_2","unstructured":"Hanchao Liu Wenyuan Xue Yifei Chen Dapeng Chen Xiutian Zhao Ke Wang Liping Hou Rongjun Li and Wei Peng. 2024. A survey on hallucination in large vision-language models. arXiv:2402.00253. Retrieved from https:\/\/arxiv.org\/abs\/2402.00253"},{"key":"e_1_3_2_94_2","doi-asserted-by":"publisher","DOI":"10.1136\/bmj.m3164"},{"key":"e_1_3_2_95_2","unstructured":"Zelong Liu Alexander Zhou Arnold Yang Alara Yilmaz Maxwell Yoo Mikey Sullivan Catherine Zhang James Grant Daiqing Li Zahi A. Fayad et al. 2023. RadImageGAN\u2014A multi-modal dataset-scale generative AI for medical imaging. arXiv:2312.05953. Retrieved from https:\/\/arxiv.org\/abs\/2312.05953"},{"key":"e_1_3_2_96_2","unstructured":"Shayne Longpre Stella Biderman Alon Albalak Hailey Schoelkopf Daniel McDuff Sayash Kapoor Kevin Klyman Kyle Lo Gabriel Ilharco Nay San et al. 2024. The Responsible foundation model development cheatsheet: A review of tools & resources. arXiv:2406.16746. Retrieved from https:\/\/arxiv.org\/abs\/2406.16746"},{"key":"e_1_3_2_97_2","unstructured":"Shayne Longpre Nikhil Singh Manuel Cherep Kushagra Tiwary Joanna Materzynska William Brannon Robert Mahari Manan Dey Mohammed Hamdy Nayan Saxena et al. 2024. Bridging the data provenance gap across text speech and video. arXiv:2412.17847. Retrieved from https:\/\/arxiv.org\/abs\/2412.17847."},{"key":"e_1_3_2_98_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41591-024-02856-4"},{"key":"e_1_3_2_99_2","doi-asserted-by":"crossref","unstructured":"Yan Luo Min Shi Muhammad Osama Khan Muhammad Muneeb Afzal Hao Huang Shuaihang Yuan Yu Tian Luo Song Ava Kouhana Tobias Elze et al. 2024. FairCLIP: Harnessing fairness in vision-language learning. arXiv:2403.19949. Retrieved from https:\/\/arxiv.org\/abs\/2403.19949","DOI":"10.1109\/CVPR52733.2024.01168"},{"key":"e_1_3_2_100_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41467-024-44824-z"},{"key":"e_1_3_2_101_2","first-page":"26230","article-title":"On the tradeoff between robustness and fairness","volume":"35","author":"Ma Xinsong","year":"2022","unstructured":"Xinsong Ma, Zekai Wang, and Weiwei Liu. 2022. On the tradeoff between robustness and fairness. In Advances in Neural Information Processing Systems, Vol. 35, 26230\u201326241.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_102_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-87240-3_37"},{"key":"e_1_3_2_103_2","doi-asserted-by":"crossref","unstructured":"Potsawee Manakul Adian Liusie and Mark J. F. Gales. 2023. SelfCheckGPT: Zero-resource black-box hallucination detection for generative large language models. arXiv:2303.08896. Retrieved from https:\/\/arxiv.org\/abs\/2303.08896","DOI":"10.18653\/v1\/2023.emnlp-main.557"},{"key":"e_1_3_2_104_2","first-page":"18445","article-title":"Ensuring fairness beyond the training data","volume":"33","author":"Mandal Debmalya","year":"2020","unstructured":"Debmalya Mandal, Samuel Deng, Suman Jana, Jeannette Wing, and Daniel J. Hsu. 2020. Ensuring fairness beyond the training data. In Advances in Neural Information Processing Systems, Vol. 33, 18445\u201318456.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_105_2","doi-asserted-by":"publisher","DOI":"10.1377\/hlthaff.2024.01003"},{"key":"e_1_3_2_106_2","doi-asserted-by":"publisher","DOI":"10.1016\/S2589-7500(20)30065-0"},{"key":"e_1_3_2_107_2","doi-asserted-by":"publisher","DOI":"10.1145\/3457607"},{"key":"e_1_3_2_108_2","doi-asserted-by":"publisher","DOI":"10.1148\/ryai.210315"},{"key":"e_1_3_2_109_2","volume-title":"Proceedings of the 39th International Conference on Machine Learning","author":"Mindermann S\u00f6ren","year":"2022","unstructured":"S\u00f6ren Mindermann, Jan M. Brauner, Muhammed T. Razzak, Mrinank Sharma, Andreas Kirsch, Winnie Xu, Benedikt H\u00f6ltgen, Aidan N. Gomez, Adrien Morisot, Sebastian Farquhar, et al. 2022. Prioritized training on points that are learnable, worth learning, and not yet learnt. In Proceedings of the 39th International Conference on Machine Learning. PMLR."},{"key":"e_1_3_2_110_2","doi-asserted-by":"publisher","DOI":"10.1001\/jama.2023.9651"},{"key":"e_1_3_2_111_2","doi-asserted-by":"publisher","DOI":"10.1145\/3287560.3287596"},{"key":"e_1_3_2_112_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41586-023-05881-4"},{"key":"e_1_3_2_113_2","doi-asserted-by":"crossref","unstructured":"Stefania L. Moroianu Christian Bluethgen Pierre Chambon Mehdi Cherti Jean-Benoit Delbrouck Magdalini Paschali Brandon Price Judy Gichoya Jenia Jitsev Curtis P. Langlotz et al. 2025. Improving performance robustness and fairness of radiographic AI models with finely-controllable synthetic data. arXiv:2508. Retrieved from https:\/\/arxiv.org\/abs\/2508.16783","DOI":"10.21203\/rs.3.rs-7687810\/v1"},{"key":"e_1_3_2_114_2","doi-asserted-by":"publisher","DOI":"10.1007\/s11704-024-3587-1"},{"key":"e_1_3_2_115_2","doi-asserted-by":"publisher","DOI":"10.1371\/journal.pdig.0000454"},{"key":"e_1_3_2_116_2","doi-asserted-by":"publisher","DOI":"10.1377\/hlthaff.2024.00842"},{"key":"e_1_3_2_117_2","doi-asserted-by":"publisher","DOI":"10.1161\/CIRCEP.119.007988"},{"key":"e_1_3_2_118_2","unstructured":"Ziad Obermeyer Rebecca Nissan Michael Stern Stephanie Eaneff Emily Joy Bembeneck and Sendhil Mullainathan. [n.d.]. Algorithmic Bias Playbook. Center for Applied AI at Chicago Booth."},{"key":"e_1_3_2_119_2","unstructured":"Maxime Oquab Timoth\u00e9e Darcet Th\u00e9o Moutakanni Huy Vo Marc Szafraniec Vasil Khalidov Pierre Fernandez Daniel Haziza Francisco Massa Alaaeldin El-Nouby et al. 2024. DINOv2: Learning robust visual features without supervision. arXiv:2304.07193. Retrieved from https:\/\/arxiv.org\/abs\/2304.07193"},{"key":"e_1_3_2_120_2","doi-asserted-by":"crossref","unstructured":"G\u00f6khan \u00d6zbulak Oscar Jimenez-del-Toro Ma\u00edra Fatoretto Lilian Berton and Andr\u00e9 Anjos. 2025. A multi-objective evaluation framework for analyzing utility-fairness trade-offs in machine learning systems. arXiv:2503.11120. Retrieved from https:\/\/arxiv.org\/abs\/2503.11120","DOI":"10.59275\/j.melba.2025-ab9a"},{"key":"e_1_3_2_121_2","doi-asserted-by":"publisher","DOI":"10.1109\/TMI.2019.2914656"},{"key":"e_1_3_2_122_2","doi-asserted-by":"crossref","unstructured":"Walter H. L. Pinaya Petru-Daniel Tudosiu Jessica Dafflon Pedro F. da Costa Virginia Fernandez Parashkev Nachev Sebastien Ourselin and M. Jorge Cardoso. 2022. Brain imaging generation with latent diffusion models. arXiv:2209.07162. Retrieved from https:\/\/arxiv.org\/abs\/2209.07162","DOI":"10.1007\/978-3-031-18576-2_12"},{"key":"e_1_3_2_123_2","first-page":"5680","article-title":"On fairness and calibration","volume":"30","author":"Pleiss Geoff","year":"2017","unstructured":"Geoff Pleiss, Manish Raghavan, Felix Wu, Jon Kleinberg, and Kilian Q. Weinberger. 2017. On fairness and calibration. In Advances in Neural Information Processing Systems, Vol. 30, 5680\u20135689.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_124_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-72787-011"},{"key":"e_1_3_2_125_2","unstructured":"Alec Radford Jong Wook Kim Chris Hallacy Aditya Ramesh Gabriel Goh Sandhini Agarwal Girish Sastry Amanda Askell Pamela Mishkin Jack Clark et al. 2021. Learning transferable visual models from natural language supervision. arXiv:2103.00020. Retrieved from https:\/\/arxiv.org\/abs\/2103.00020"},{"key":"e_1_3_2_126_2","doi-asserted-by":"publisher","DOI":"10.7326\/M18-1990"},{"key":"e_1_3_2_127_2","unstructured":"Pranav Rajpurkar Jeremy Irvin Aarti Bagul Daisy Ding Tony Duan Hershel Mehta Brandon Yang Kaylie Zhu Dillon Laird Robyn L. Ball et al. 2018. MURA: Large dataset for abnormality detection in musculoskeletal radiographs. arXiv:1712.06957. Retrieved from https:\/\/arxiv.org\/abs\/1712.06957"},{"key":"e_1_3_2_128_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41597-022-01608-8"},{"key":"e_1_3_2_129_2","unstructured":"Eduardo Pontes Reis Felipe Nascimento Mateus Aranha Fernando Mainetti Secol Birajara Machado Marcelo Felix Anouk Stein and Edson Amaro. 2020. Brain hemorrhage extended (BHX): Bounding box extrapolation from thick to thin slice CT images. PhysioNet 101 (2020) e215\u2013e220."},{"key":"e_1_3_2_130_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41467-022-32186-3"},{"key":"e_1_3_2_131_2","unstructured":"Tom Schaul John Quan Ioannis Antonoglou and David Silver. 2015. Prioritized experience replay. arXiv:1511.05952. Retrieved from https:\/\/arxiv.org\/abs\/1511.05952"},{"key":"e_1_3_2_132_2","unstructured":"Jessica Schrouff Natalie Harris Oluwasanmi Koyejo Ibrahim Alabdulmohsin Eva Schnider Krista Opsahl-Ong Alex Brown Subhrajit Roy Diana Mincu Christina Chen et al. 2023. Diagnosing failures of fairness transfer across distribution shift in real-world medical settings. arXiv:2202.01034. Retrieved from https:\/\/arxiv.org\/abs\/2202.01034"},{"key":"e_1_3_2_133_2","unstructured":"Andrew Sellergren Sahar Kazemzadeh Tiam Jaroensri Atilla Kiraly Madeleine Traverse Timo Kohlberger Shawn Xu Fayaz Jamil C\u00edan Hughes Charles Lau et al. 2025. MedGemma technical report. arXiv:2507.05201. Retrieved from https:\/\/arxiv.org\/abs\/2507.05201"},{"key":"e_1_3_2_134_2","unstructured":"Congzhen Shi Ryan Rezai Jiaxi Yang Qi Dou and Xiaoxiao Li. 2024. A survey on trustworthiness in foundation models for medical image analysis. arXiv:2407.15851. Retrieved from https:\/\/arxiv.org\/abs\/2407.15851"},{"key":"e_1_3_2_135_2","doi-asserted-by":"crossref","unstructured":"Adi Simhi Itay Itzhak Fazl Barez Gabriel Stanovsky and Yonatan Belinkov. 2025. Trust me I\u2019m wrong: LLMs hallucinate with certainty despite knowing the answer. arXiv:2502.12964. Retrieved from https:\/\/arxiv.org\/abs\/2502.12964","DOI":"10.18653\/v1\/2025.findings-emnlp.792"},{"key":"e_1_3_2_136_2","unstructured":"Yang Song and Prafulla Dhariwal. 2023. Improved techniques for training consistency models. arXiv:2310.14189. Retrieved from https:\/\/arxiv.org\/abs\/2310.14189"},{"key":"e_1_3_2_137_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.ebiom.2024.105501"},{"key":"e_1_3_2_138_2","doi-asserted-by":"publisher","DOI":"10.1109\/TMI.2015.2487997"},{"key":"e_1_3_2_139_2","unstructured":"L. C. M. Team Lo\u00efc Barrault Paul-Ambroise Duquenne Maha Elbayad Artyom Kozhevnikov Belen Alastruey Pierre Andrews Mariano Coria Guillaume Couairon Marta R. Costa-Juss\u00e0 et al. 2024. Large concept models: Language modeling in a sentence representation space. arXiv:2412.08821. Retrieved from https:\/\/arxiv.org\/abs\/2412.08821"},{"key":"e_1_3_2_140_2","doi-asserted-by":"publisher","DOI":"10.1148\/ryai.240300"},{"key":"e_1_3_2_141_2","unstructured":"Cuong Tran Ferdinando Fioretto Jung-Eun Kim and Rakshit Naidu. 2022. Pruning has a disparate impact on model accuracy. arXiv:2205.13574. Retrieved from https:\/\/arxiv.org\/abs\/2205.13574"},{"key":"e_1_3_2_142_2","unstructured":"Michael Tschannen Alexey Gritsenko Xiao Wang Muhammad Ferjad Naeem Ibrahim Alabdulmohsin Nikhil Parthasarathy Talfan Evans Lucas Beyer Ye Xia Basil Mustafa et al. 2025. SigLIP 2: Multilingual vision-language encoders with improved semantic understanding localization and dense features. arXiv:2502.14786. Retrieved from https:\/\/arxiv.org\/abs\/2502.14786"},{"key":"e_1_3_2_143_2","unstructured":"Tao Tu Shekoofeh Azizi Danny Driess Mike Schaekermann Mohamed Amin Pi-Chuan Chang Andrew Carroll Chuck Lau Ryutaro Tanno Ira Ktena et al. 2023. Towards generalist biomedical AI. arXiv230714334. Retrieved from https:\/\/arxiv.org\/abs\/2307.14334"},{"key":"e_1_3_2_144_2","volume-title":"National Archives and Records Administration\u2014An Act Entitled the Patient Protection and Affordable Care Act","author":"Office of the Federal Register, United States","year":"2010","unstructured":"Office of the Federal Register, United States. 2010. National Archives and Records Administration\u2014An Act Entitled the Patient Protection and Affordable Care Act. Office of the Federal Register, United States."},{"key":"e_1_3_2_145_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41591-024-02885-z"},{"key":"e_1_3_2_146_2","doi-asserted-by":"publisher","DOI":"10.1136\/bmj-2022-070904"},{"key":"e_1_3_2_147_2","unstructured":"Huy V. Vo Vasil Khalidov Timoth\u00e9e Darcet Th\u00e9o Moutakanni Nikita Smetanin Marc Szafraniec Hugo Touvron Camille Couprie Maxime Oquab Armand Joulin et al. 2024. Automatic data curation for self-supervised learning: A clustering-based approach. arXiv:2405.15613. Retrieved from https:\/\/arxiv.org\/abs\/2405.15613"},{"key":"e_1_3_2_148_2","unstructured":"Xiyao Wang Jiuhai Chen Zhaoyang Wang Yuhang Zhou Yiyang Zhou Huaxiu Yao Tianyi Zhou Tom Goldstein Parminder Bhatia Furong Huang et al. 2024. Enhancing visual-language modality alignment in large vision language models via self-improvement. arXiv:2405.15973. Retrieved from https:\/\/arxiv.org\/abs\/2405.15973"},{"key":"e_1_3_2_149_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.369"},{"key":"e_1_3_2_150_2","doi-asserted-by":"crossref","unstructured":"Zifeng Wang Zhenbang Wu Dinesh Agarwal and Jimeng Sun. 2022. MedCLIP: Contrastive learning from unpaired medical images and text. arXiv:2210.10163. Retrieved from https:\/\/arxiv.org\/abs\/2210.10163","DOI":"10.18653\/v1\/2022.emnlp-main.256"},{"key":"e_1_3_2_151_2","unstructured":"Susan Wei and Marc Niethammer. 2021. The fairness-accuracy pareto front. arXiv:2008.10797. Retrieved from https:\/\/arxiv.org\/abs\/2008.10797"},{"key":"e_1_3_2_152_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41591-023-02293-9"},{"key":"e_1_3_2_153_2","doi-asserted-by":"publisher","DOI":"10.1111\/1475-6773.13222"},{"key":"e_1_3_2_154_2","doi-asserted-by":"publisher","DOI":"10.1111\/epi.16320"},{"key":"e_1_3_2_155_2","unstructured":"World Health Organization. 2010. A Conceptual Framework for Action on the Social Determinants of Health. World Health Organization."},{"key":"e_1_3_2_156_2","unstructured":"Chaoyi Wu Xiaoman Zhang Ya Zhang Yanfeng Wang and Weidi Xie. 2023. Towards generalist foundation model for radiology by leveraging web-scale 2D&3D medical data. arXiv:2308.02463. Retrieved from https:\/\/arxiv.org\/abs\/2308.02463"},{"key":"e_1_3_2_157_2","first-page":"140334","article-title":"CARES: A comprehensive benchmark of trustworthiness in medical vision language models","volume":"37","author":"Xia Peng","year":"2024","unstructured":"Peng Xia, Ze Chen, Juanxi Tian, Yangrui Gong, Ruibo Hou, Yue Xu, Zhenbang Wu, Zhiyuan Fan, Yiyang Zhou, Kangyu Zhu, et al. 2024. CARES: A comprehensive benchmark of trustworthiness in medical vision language models. In Advances in Neural Information Processing Systems, Vol. 37, 140334\u2013140365","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_158_2","unstructured":"Yijun Xiao and William Yang Wang. 2021. On hallucination and predictive uncertainty in conditional language generation. arXiv:2103.15025. Retrieved from https:\/\/arxiv.org\/abs\/2103.15025"},{"key":"e_1_3_2_159_2","unstructured":"Yunfei Xie Ce Zhou Lang Gao Juncheng Wu Xianhang Li Hong-Yu Zhou Sheng Liu Lei Xing James Zou Cihang Xie et al. 2024. MedTrinity-25M: A large-scale multimodal dataset with multigranular annotations for medicine. arXiv:2408. Retrieved from https:\/\/arxiv.org\/abs\/2408.02900"},{"key":"e_1_3_2_160_2","unstructured":"Han Xu Xiaorui Liu Yaxin Li Anil K. Jain and Jiliang Tang. 2021. To be robust or to be fair: Towards fairness in adversarial training. arXiv:2010.06121. Retrieved from https:\/\/arxiv.org\/abs\/2010.06121"},{"key":"e_1_3_2_161_2","volume-title":"Proceedings of the 12th International Conference on Learning Representations","author":"Xu Hu","year":"2023","unstructured":"Hu Xu, Saining Xie, Xiaoqing Tan, Po-Yao Huang, Russell Howes, Vasu Sharma, Shang-Wen Li, Gargi Ghosh, Luke Zettlemoyer, and Christoph Feichtenhofer. 2023. Demystifying CLIP data. In Proceedings of the 12th International Conference on Learning Representations."},{"key":"e_1_3_2_162_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41746-024-01276-5"},{"key":"e_1_3_2_163_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41591-024-03113-4"},{"key":"e_1_3_2_164_2","unstructured":"Muhammad Bilal Zafar Isabel Valera Manuel Gomez Rodriguez and Krishna P. Gummadi. 2017. Fairness constraints: Mechanisms for fair classification. arXiv:1507.05259. Retrieved from https:\/\/arxiv.org\/abs\/1507.05259"},{"key":"e_1_3_2_165_2","unstructured":"Anna Zawacki. 2020. SIIM-ISIC Melanoma Classification. Retrieved from https:\/\/kaggle.com\/siim-isic-melanoma-classification"},{"key":"e_1_3_2_166_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV51070.2023.01100"},{"key":"e_1_3_2_167_2","doi-asserted-by":"crossref","unstructured":"Haoran Zhang Natalie Dullerud Karsten Roth Lauren Oakden-Rayner Stephen Robert Pfohl and Marzyeh Ghassemi. 2022. Improving the fairness of chest X-ray classifiers. arXiv:2203.12609. Retrieved from https:\/\/arxiv.org\/abs\/2203.12609","DOI":"10.1109\/EMBC48229.2022.9871784"},{"key":"e_1_3_2_168_2","unstructured":"Ruichen Zhang Yuguang Yao Zhen Tan Zhiming Li Pan Wang Huan Liu Jingtong Hu Sijia Liu and Tianlong Chen. 2024. FairSkin: Fair diffusion for skin disease image generation. arXiv:2410.22551. Retrieved from https:\/\/arxiv.org\/abs\/2410.22551"},{"key":"e_1_3_2_169_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.media.2023.102996"},{"key":"e_1_3_2_170_2","doi-asserted-by":"crossref","unstructured":"Sheng Zhang Yanbo Xu Naoto Usuyama Hanwen Xu Jaspreet Bagga Robert Tinn Sam Preston Rajesh Rao Mu Wei Naveen Valluri et al. 2025. BiomedCLIP: A multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs. arXiv:2303.00915. Retrieved from https:\/\/arxiv.org\/abs\/2303.00915","DOI":"10.1056\/AIoa2400640"},{"key":"e_1_3_2_171_2","unstructured":"Xiaoman Zhang Chaoyi Wu Ziheng Zhao Weixiong Lin Ya Zhang Yanfeng Wang and Weidi Xie. 2023. PMC-VQA: Visual instruction tuning for medical visual question answering. arXiv:2305.10415. Retrieved from https:\/\/arxiv.org\/abs\/2305.10415"},{"key":"e_1_3_2_172_2","unstructured":"Han Zhao and Geoffrey J. Gordon. 2022. Inherent tradeoffs in learning fair representations. arXiv:1906.08386. Retrieved from https:\/\/arxiv.org\/abs\/1906.08386"},{"key":"e_1_3_2_173_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41592-024-02499-w"},{"key":"e_1_3_2_174_2","doi-asserted-by":"publisher","DOI":"10.1038\/s41586-023-06555-x"},{"key":"e_1_3_2_175_2","unstructured":"Yuyin Zhou Shih-Cheng Huang Jason Alan Fries Alaa Youssef Timothy J. Amrhein Marcello Chang Imon Banerjee Daniel Rubin Lei Xing Nigam Shah et al. 2021. RadFusion: Benchmarking performance and fairness for multimodal pulmonary embolism detection from CT and EHR. arXiv:2111.11665. Retrieved from https:\/\/arxiv.org\/abs\/2111.11665"},{"key":"e_1_3_2_176_2","unstructured":"Yongshuo Zong Yongxin Yang and Timothy Hospedales. 2023. MEDFAIR: Benchmarking fairness for medical imaging. arXiv:2210.01725. Retrieved from https:\/\/arxiv.org\/abs\/2210.01725"},{"key":"e_1_3_2_177_2","doi-asserted-by":"publisher","DOI":"10.1126\/science.adh4260"},{"key":"e_1_3_2_178_2","doi-asserted-by":"publisher","DOI":"10.1145\/3461702.3462617"}],"container-title":["ACM Transactions on Computing for Healthcare"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3793542","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,3,17]],"date-time":"2026-03-17T15:10:27Z","timestamp":1773760227000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3793542"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,3,17]]},"references-count":177,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2026,4,30]]}},"alternative-id":["10.1145\/3793542"],"URL":"https:\/\/doi.org\/10.1145\/3793542","relation":{},"ISSN":["2691-1957","2637-8051"],"issn-type":[{"value":"2691-1957","type":"print"},{"value":"2637-8051","type":"electronic"}],"subject":[],"published":{"date-parts":[[2026,3,17]]},"assertion":[{"value":"2025-03-28","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2026-01-13","order":2,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2026-03-17","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}