{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,8,28]],"date-time":"2026-08-28T05:16:02Z","timestamp":1787894162849,"version":"build-2784847793"},"publisher-location":"New York, NY, USA","reference-count":80,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,7,26]],"date-time":"2022-07-26T00:00:00Z","timestamp":1658793600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by-nc-sa\/4.0\/"}],"funder":[{"name":"National Institute of Standards and Technology (NIST)","award":["60NANB20D212T"],"award-info":[{"award-number":["60NANB20D212T"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,7,26]]},"DOI":"10.1145\/3514094.3534136","type":"proceedings-article","created":{"date-parts":[[2022,7,27]],"date-time":"2022-07-27T22:25:13Z","timestamp":1658960713000},"page":"800-812","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":28,"title":["American == White in Multimodal Language-and-Image AI"],"prefix":"10.1145","author":[{"given":"Robert","family":"Wolfe","sequence":"first","affiliation":[{"name":"University of Washington, Seattle, WA, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Aylin","family":"Caliskan","sequence":"additional","affiliation":[{"name":"University of Washington, Seattle, WA, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2022,7,27]]},"reference":[{"key":"e_1_3_2_2_1_1","volume-title":"Jong Wook Kim, and Miles Brundage","author":"Agarwal Sandhini","year":"2021","unstructured":"Sandhini Agarwal , Gretchen Krueger , Jack Clark , Alec Radford , Jong Wook Kim, and Miles Brundage . 2021 . Evaluating CLIP: Towards Characterization of Broader Capabilities and Downstream Implications . arXiv preprint arXiv:2108.02818 (2021). Sandhini Agarwal, Gretchen Krueger, Jack Clark, Alec Radford, Jong Wook Kim, and Miles Brundage. 2021. Evaluating CLIP: Towards Characterization of Broader Capabilities and Downstream Implications. arXiv preprint arXiv:2108.02818 (2021)."},{"key":"e_1_3_2_2_2_1","doi-asserted-by":"publisher","DOI":"10.1037\/a0031547"},{"key":"e_1_3_2_2_3_1","volume-title":"Vinay Uday Prabhu, and Emmanuel Kahembwe","author":"Birhane Abeba","year":"2021","unstructured":"Abeba Birhane , Vinay Uday Prabhu, and Emmanuel Kahembwe . 2021 . Multimodal datasets: misogyny, pornography, and malignant stereotypes. arXiv preprint arXiv:2110.01963 (2021). Abeba Birhane, Vinay Uday Prabhu, and Emmanuel Kahembwe. 2021. Multimodal datasets: misogyny, pornography, and malignant stereotypes. arXiv preprint arXiv:2110.01963 (2021)."},{"key":"e_1_3_2_2_4_1","volume-title":"International Conference on Machine Learning. PMLR, 803--811","author":"Brunet Marc-Etienne","year":"2019","unstructured":"Marc-Etienne Brunet , Colleen Alkalay-Houlihan , Ashton Anderson , and Richard Zemel . 2019 . Understanding the origins of bias in word embeddings . In International Conference on Machine Learning. PMLR, 803--811 . Marc-Etienne Brunet, Colleen Alkalay-Houlihan, Ashton Anderson, and Richard Zemel. 2019. Understanding the origins of bias in word embeddings. In International Conference on Machine Learning. PMLR, 803--811."},{"key":"e_1_3_2_2_5_1","volume-title":"Conference on fairness, accountability and transparency. PMLR, 77--91","author":"Buolamwini Joy","year":"2018","unstructured":"Joy Buolamwini and Timnit Gebru . 2018 . Gender shades: Intersectional accuracy disparities in commercial gender classification . In Conference on fairness, accountability and transparency. PMLR, 77--91 . Joy Buolamwini and Timnit Gebru. 2018. Gender shades: Intersectional accuracy disparities in commercial gender classification. In Conference on fairness, accountability and transparency. PMLR, 77--91."},{"key":"e_1_3_2_2_6_1","volume-title":"Proceedings of the 2022 AAAI\/ACM Conference on AI, Ethics, and Society.","author":"Caliskan Aylin","unstructured":"Aylin Caliskan , Pimparkar Parth Ajay , Tessa Charlesworth , Robert Wolfe , and Mahzarin R. Banaji . 2022. Gender Bias in Word Embeddings: A Comprehensive Analysis of Frequency, Syntax, and Semantics . In Proceedings of the 2022 AAAI\/ACM Conference on AI, Ethics, and Society. Aylin Caliskan, Pimparkar Parth Ajay, Tessa Charlesworth, Robert Wolfe, and Mahzarin R. Banaji. 2022. Gender Bias in Word Embeddings: A Comprehensive Analysis of Frequency, Syntax, and Semantics. In Proceedings of the 2022 AAAI\/ACM Conference on AI, Ethics, and Society."},{"key":"e_1_3_2_2_7_1","volume-title":"Science","volume":"356","author":"Caliskan Aylin","year":"2017","unstructured":"Aylin Caliskan , Joanna J Bryson , and Arvind Narayanan . 2017 . Semantics derived automatically from language corpora contain human-like biases . Science , Vol. 356 , 6334 (2017), 183--186. Aylin Caliskan, Joanna J Bryson, and Arvind Narayanan. 2017. Semantics derived automatically from language corpora contain human-like biases. Science, Vol. 356, 6334 (2017), 183--186."},{"key":"e_1_3_2_2_8_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.00356"},{"key":"e_1_3_2_2_9_1","volume-title":"International Conference on Machine Learning. PMLR, 1691--1703","author":"Chen Mark","year":"2020","unstructured":"Mark Chen , Alec Radford , Rewon Child , Jeffrey Wu , Heewoo Jun , David Luan , and Ilya Sutskever . 2020 . Generative pretraining from pixels . In International Conference on Machine Learning. PMLR, 1691--1703 . Mark Chen, Alec Radford, Rewon Child, Jeffrey Wu, Heewoo Jun, David Luan, and Ilya Sutskever. 2020. Generative pretraining from pixels. In International Conference on Machine Learning. PMLR, 1691--1703."},{"key":"e_1_3_2_2_10_1","volume-title":"Statistical power analysis. Current directions in psychological science","author":"Cohen Jacob","year":"1992","unstructured":"Jacob Cohen . 1992. Statistical power analysis. Current directions in psychological science , Vol. 1 , 3 ( 1992 ), 98--101. Jacob Cohen. 1992. Statistical power analysis. Current directions in psychological science, Vol. 1, 3 (1992), 98--101."},{"key":"e_1_3_2_2_11_1","volume-title":"VQGAN-CLIP: Open Domain Image Generation and Editing with Natural Language Guidance. arXiv preprint arXiv:2204.08583","author":"Crowson Katherine","year":"2022","unstructured":"Katherine Crowson , Stella Biderman , Daniel Kornis , Dashiell Stander , Eric Hallahan , Louis Castricato , and Edward Raff . 2022. VQGAN-CLIP: Open Domain Image Generation and Editing with Natural Language Guidance. arXiv preprint arXiv:2204.08583 ( 2022 ). Katherine Crowson, Stella Biderman, Daniel Kornis, Dashiell Stander, Eric Hallahan, Louis Castricato, and Edward Raff. 2022. VQGAN-CLIP: Open Domain Image Generation and Editing with Natural Language Guidance. arXiv preprint arXiv:2204.08583 (2022)."},{"key":"e_1_3_2_2_12_1","doi-asserted-by":"publisher","DOI":"10.1177\/1948550614546355"},{"key":"e_1_3_2_2_13_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2009.5206848"},{"key":"e_1_3_2_2_14_1","first-page":"19","volume-title":"Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies","volume":"1","author":"Devlin Jacob","year":"2019","unstructured":"Jacob Devlin , Ming-Wei Chang , Kenton Lee , and Kristina Toutanova . 2019 . BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding . In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies , Volume 1 (Long and Short Papers). Association for Computational Linguistics, Minneapolis, Minnesota, 4171--4186. https:\/\/doi.org\/10. 18653\/v1\/N 19 - 1423 10.18653\/v1 Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019. BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers). Association for Computational Linguistics, Minneapolis, Minnesota, 4171--4186. https:\/\/doi.org\/10.18653\/v1\/N19-1423"},{"key":"e_1_3_2_2_15_1","doi-asserted-by":"publisher","DOI":"10.1037\/0022-3514.88.3.447"},{"key":"e_1_3_2_2_16_1","doi-asserted-by":"publisher","DOI":"10.1177\/1368430219887440"},{"key":"e_1_3_2_2_17_1","volume-title":"Documenting the english colossal clean crawled corpus. arXiv preprint arXiv:2104.08758","author":"Dodge Jesse","year":"2021","unstructured":"Jesse Dodge , Maarten Sap , Ana Marasovic , William Agnew , Gabriel Ilharco , Dirk Groeneveld , and Matt Gardner . 2021. Documenting the english colossal clean crawled corpus. arXiv preprint arXiv:2104.08758 ( 2021 ). Jesse Dodge, Maarten Sap, Ana Marasovic, William Agnew, Gabriel Ilharco, Dirk Groeneveld, and Matt Gardner. 2021. Documenting the english colossal clean crawled corpus. arXiv preprint arXiv:2104.08758 (2021)."},{"key":"e_1_3_2_2_18_1","volume-title":"International Conference on Learning Representations.","author":"Dosovitskiy Alexey","year":"2020","unstructured":"Alexey Dosovitskiy , Lucas Beyer , Alexander Kolesnikov , Dirk Weissenborn , Xiaohua Zhai , Thomas Unterthiner , Mostafa Dehghani , Matthias Minderer , Georg Heigold , Sylvain Gelly , 2020 . An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale . In International Conference on Learning Representations. Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al. 2020. An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale. In International Conference on Learning Representations."},{"key":"e_1_3_2_2_19_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.01268"},{"key":"e_1_3_2_2_20_1","volume-title":"Devise: A deep visual-semantic embedding model.","author":"Frome Andrea","year":"2013","unstructured":"Andrea Frome , Greg Corrado , Jonathon Shlens , Samy Bengio , Jeffrey Dean , Marc'Aurelio Ranzato , and Tomas Mikolov . 2013 . Devise: A deep visual-semantic embedding model. (2013). Andrea Frome, Greg Corrado, Jonathon Shlens, Samy Bengio, Jeffrey Dean, Marc'Aurelio Ranzato, and Tomas Mikolov. 2013. Devise: A deep visual-semantic embedding model. (2013)."},{"key":"e_1_3_2_2_21_1","doi-asserted-by":"publisher","DOI":"10.23915\/distill.00030"},{"key":"e_1_3_2_2_22_1","doi-asserted-by":"publisher","DOI":"10.1037\/0022-3514.74.6.1464"},{"key":"e_1_3_2_2_23_1","volume-title":"Citizenship and ethnicity: the growth and development of a democratic multiethnic institution. Number 128","author":"Gross Feliks","unstructured":"Feliks Gross . 1999. Citizenship and ethnicity: the growth and development of a democratic multiethnic institution. Number 128 . Greenwood Publishing Group . Feliks Gross. 1999. Citizenship and ethnicity: the growth and development of a democratic multiethnic institution. Number 128. Greenwood Publishing Group."},{"key":"e_1_3_2_2_24_1","volume-title":"Open-vocabulary Object Detection via Vision and Language Knowledge Distillation. arXiv preprint arXiv:2104.13921","author":"Gu Xiuye","year":"2021","unstructured":"Xiuye Gu , Tsung-Yi Lin , Weicheng Kuo , and Yin Cui . 2021. Open-vocabulary Object Detection via Vision and Language Knowledge Distillation. arXiv preprint arXiv:2104.13921 , Vol. 2 ( 2021 ). Xiuye Gu, Tsung-Yi Lin, Weicheng Kuo, and Yin Cui. 2021. Open-vocabulary Object Detection via Vision and Language Knowledge Distillation. arXiv preprint arXiv:2104.13921, Vol. 2 (2021)."},{"key":"e_1_3_2_2_25_1","doi-asserted-by":"publisher","DOI":"10.1145\/3461702.3462536"},{"key":"e_1_3_2_2_26_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.90"},{"key":"e_1_3_2_2_27_1","doi-asserted-by":"publisher","DOI":"10.1111\/pops.12189"},{"key":"e_1_3_2_2_28_1","doi-asserted-by":"publisher","DOI":"10.1521\/jscp.2011.30.2.133"},{"key":"e_1_3_2_2_29_1","volume-title":"Scaling Up Visual and Vision-Language Representation Learning With Noisy Text Supervision. arXiv e-prints","author":"Jia Chao","year":"2021","unstructured":"Chao Jia , Yinfei Yang , Ye Xia , Yi-Ting Chen , Zarana Parekh , Hieu Pham , Quoc V Le , Yunhsuan Sung , Zhen Li , and Tom Duerig . 2021. Scaling Up Visual and Vision-Language Representation Learning With Noisy Text Supervision. arXiv e-prints ( 2021 ), arXiv-2102. Chao Jia, Yinfei Yang, Ye Xia, Yi-Ting Chen, Zarana Parekh, Hieu Pham, Quoc V Le, Yunhsuan Sung, Zhen Li, and Tom Duerig. 2021. Scaling Up Visual and Vision-Language Representation Learning With Noisy Text Supervision. arXiv e-prints (2021), arXiv-2102."},{"key":"e_1_3_2_2_30_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2020.acl-main.405"},{"key":"e_1_3_2_2_31_1","doi-asserted-by":"publisher","DOI":"10.1145\/3461702.3462609"},{"key":"e_1_3_2_2_32_1","doi-asserted-by":"publisher","DOI":"10.1177\/0003122419877135"},{"key":"e_1_3_2_2_33_1","volume-title":"Visual Genome: Connecting Language and Vision Using Crowdsourced Dense Image Annotations. https:\/\/arxiv.org\/abs\/1602.07332","author":"Krishna Ranjay","year":"2016","unstructured":"Ranjay Krishna , Yuke Zhu , Oliver Groth , Justin Johnson , Kenji Hata , Joshua Kravitz , Stephanie Chen , Yannis Kalantidis , Li-Jia Li , David A Shamma , Michael Bernstein , and Li Fei-Fei . 2016 . Visual Genome: Connecting Language and Vision Using Crowdsourced Dense Image Annotations. https:\/\/arxiv.org\/abs\/1602.07332 Ranjay Krishna, Yuke Zhu, Oliver Groth, Justin Johnson, Kenji Hata, Joshua Kravitz, Stephanie Chen, Yannis Kalantidis, Li-Jia Li, David A Shamma, Michael Bernstein, and Li Fei-Fei. 2016. Visual Genome: Connecting Language and Vision Using Crowdsourced Dense Image Annotations. https:\/\/arxiv.org\/abs\/1602.07332"},{"key":"e_1_3_2_2_34_1","volume-title":"Proceedings of the IEEE International Conference on Computer Vision. 4183--4192","author":"Li Ang","unstructured":"Ang Li , Allan Jabri , Armand Joulin , and Laurens van der Maaten. 2017. Learning visual n-grams from web data . In Proceedings of the IEEE International Conference on Computer Vision. 4183--4192 . Ang Li, Allan Jabri, Armand Joulin, and Laurens van der Maaten. 2017. Learning visual n-grams from web data. In Proceedings of the IEEE International Conference on Computer Vision. 4183--4192."},{"key":"e_1_3_2_2_35_1","volume-title":"BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and Generation. arXiv preprint arXiv:2201.12086","author":"Li Junnan","year":"2022","unstructured":"Junnan Li , Dongxu Li , Caiming Xiong , and Steven Hoi . 2022 . BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and Generation. arXiv preprint arXiv:2201.12086 (2022). Junnan Li, Dongxu Li, Caiming Xiong, and Steven Hoi. 2022. BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and Generation. arXiv preprint arXiv:2201.12086 (2022)."},{"key":"e_1_3_2_2_36_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-10602-1_48"},{"key":"e_1_3_2_2_37_1","volume-title":"The Chicago face database: A free stimulus set of faces and norming data. Behavior research methods","author":"Ma Debbie S","year":"2015","unstructured":"Debbie S Ma , Joshua Correll , and Bernd Wittenbrink . 2015. The Chicago face database: A free stimulus set of faces and norming data. Behavior research methods , Vol. 47 , 4 ( 2015 ), 1122--1135. Debbie S Ma, Joshua Correll, and Bernd Wittenbrink. 2015. The Chicago face database: A free stimulus set of faces and norming data. Behavior research methods, Vol. 47, 4 (2015), 1122--1135."},{"key":"e_1_3_2_2_38_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/N19-1063"},{"key":"e_1_3_2_2_39_1","volume-title":"SLIP: Self-supervision meets Language-Image Pre-training. arXiv preprint arXiv:2112.12750","author":"Mu Norman","year":"2021","unstructured":"Norman Mu , Alexander Kirillov , David Wagner , and Saining Xie . 2021 . SLIP: Self-supervision meets Language-Image Pre-training. arXiv preprint arXiv:2112.12750 (2021). Norman Mu, Alexander Kirillov, David Wagner, and Saining Xie. 2021. SLIP: Self-supervision meets Language-Image Pre-training. arXiv preprint arXiv:2112.12750 (2021)."},{"key":"e_1_3_2_2_40_1","volume-title":"MUM: A new AI milestone for understanding information. https:\/\/blog.google\/products\/search\/introducing-mum\/","author":"Nayak Pandu","year":"2021","unstructured":"Pandu Nayak . 2021 . MUM: A new AI milestone for understanding information. https:\/\/blog.google\/products\/search\/introducing-mum\/ Pandu Nayak. 2021. MUM: A new AI milestone for understanding information. https:\/\/blog.google\/products\/search\/introducing-mum\/"},{"key":"e_1_3_2_2_41_1","volume-title":"Glide: Towards photorealistic image generation and editing with text-guided diffusion models. arXiv preprint arXiv:2112.10741","author":"Nichol Alex","year":"2021","unstructured":"Alex Nichol , Prafulla Dhariwal , Aditya Ramesh , Pranav Shyam , Pamela Mishkin , Bob McGrew , Ilya Sutskever , and Mark Chen . 2021 . Glide: Towards photorealistic image generation and editing with text-guided diffusion models. arXiv preprint arXiv:2112.10741 (2021). Alex Nichol, Prafulla Dhariwal, Aditya Ramesh, Pranav Shyam, Pamela Mishkin, Bob McGrew, Ilya Sutskever, and Mark Chen. 2021. Glide: Towards photorealistic image generation and editing with text-guided diffusion models. arXiv preprint arXiv:2112.10741 (2021)."},{"key":"e_1_3_2_2_42_1","volume-title":"Structural racism, economic opportunity and racial health disparities: Evidence from US counties. SSM-Population health","author":"O'Brien Rourke","year":"2020","unstructured":"Rourke O'Brien , Tiffany Neman , Nathan Seltzer , Linnea Evans , and Atheendar Venkataramani . 2020. Structural racism, economic opportunity and racial health disparities: Evidence from US counties. SSM-Population health , Vol. 11 ( 2020 ), 100564. Rourke O'Brien, Tiffany Neman, Nathan Seltzer, Linnea Evans, and Atheendar Venkataramani. 2020. Structural racism, economic opportunity and racial health disparities: Evidence from US counties. SSM-Population health, Vol. 11 (2020), 100564."},{"key":"e_1_3_2_2_43_1","volume-title":"Im2text: Describing images using 1 million captioned photographs. Advances in neural information processing systems","author":"Ordonez Vicente","year":"2011","unstructured":"Vicente Ordonez , Girish Kulkarni , and Tamara Berg . 2011. Im2text: Describing images using 1 million captioned photographs. Advances in neural information processing systems , Vol. 24 ( 2011 ). Vicente Ordonez, Girish Kulkarni, and Tamara Berg. 2011. Im2text: Describing images using 1 million captioned photographs. Advances in neural information processing systems, Vol. 24 (2011)."},{"key":"e_1_3_2_2_44_1","doi-asserted-by":"publisher","DOI":"10.1145\/3461702.3462561"},{"key":"e_1_3_2_2_45_1","volume-title":"Understanding the Representation and Representativeness of Age in AI Data Sets. arXiv preprint arXiv:2103.09058","author":"Park Joon Sung","year":"2021","unstructured":"Joon Sung Park , Michael S Bernstein , Robin N Brewer , Ece Kamar , and Meredith Ringel Morris . 2021. Understanding the Representation and Representativeness of Age in AI Data Sets. arXiv preprint arXiv:2103.09058 ( 2021 ). Joon Sung Park, Michael S Bernstein, Robin N Brewer, Ece Kamar, and Meredith Ringel Morris. 2021. Understanding the Representation and Representativeness of Age in AI Data Sets. arXiv preprint arXiv:2103.09058 (2021)."},{"key":"e_1_3_2_2_46_1","volume-title":"The Quest for the Perfect \"Smart Wall'. Politico","author":"Phippen J. Weston","year":"2021","unstructured":"J. Weston Phippen . 2021. \" A $10- Million Scarecrow' : The Quest for the Perfect \"Smart Wall'. Politico ( 2021 ). J. Weston Phippen. 2021. \"A $10-Million Scarecrow': The Quest for the Perfect \"Smart Wall'. Politico (2021)."},{"key":"e_1_3_2_2_47_1","volume-title":"Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al.","author":"Radford Alec","year":"2021","unstructured":"Alec Radford , Jong Wook Kim , Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al. 2021 . Learning transferable visual models from natural language supervision. arXiv preprint arXiv:2103.00020 (2021). Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al. 2021. Learning transferable visual models from natural language supervision. arXiv preprint arXiv:2103.00020 (2021)."},{"key":"e_1_3_2_2_48_1","unstructured":"Alec Radford Jeffrey Wu Rewon Child David Luan Dario Amodei Ilya Sutskever etal 2019. Language models are unsupervised multitask learners. OpenAI blog Vol. 1 8 (2019) 9.  Alec Radford Jeffrey Wu Rewon Child David Luan Dario Amodei Ilya Sutskever et al. 2019. Language models are unsupervised multitask learners. OpenAI blog Vol. 1 8 (2019) 9."},{"key":"e_1_3_2_2_49_1","volume-title":"Hierarchical Text-Conditional Image Generation with CLIP Latents. arXiv preprint arXiv:2204.06125","author":"Ramesh Aditya","year":"2022","unstructured":"Aditya Ramesh , Prafulla Dhariwal , Alex Nichol , Casey Chu , and Mark Chen . 2022. Hierarchical Text-Conditional Image Generation with CLIP Latents. arXiv preprint arXiv:2204.06125 ( 2022 ). Aditya Ramesh, Prafulla Dhariwal, Alex Nichol, Casey Chu, and Mark Chen. 2022. Hierarchical Text-Conditional Image Generation with CLIP Latents. arXiv preprint arXiv:2204.06125 (2022)."},{"key":"e_1_3_2_2_50_1","volume-title":"Zero-shot text-to-image generation. arXiv preprint arXiv:2102.12092","author":"Ramesh Aditya","year":"2021","unstructured":"Aditya Ramesh , Mikhail Pavlov , Gabriel Goh , Scott Gray , Chelsea Voss , Alec Radford , Mark Chen , and Ilya Sutskever . 2021. Zero-shot text-to-image generation. arXiv preprint arXiv:2102.12092 ( 2021 ). Aditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray, Chelsea Voss, Alec Radford, Mark Chen, and Ilya Sutskever. 2021. Zero-shot text-to-image generation. arXiv preprint arXiv:2102.12092 (2021)."},{"key":"e_1_3_2_2_51_1","volume-title":"to End Use of Facial Recognition for Identity Verification. The New York Times (Feb","author":"Rappeport Alan","year":"2022","unstructured":"Alan Rappeport and Kashmir Hill . 2022. I.R. S. to End Use of Facial Recognition for Identity Verification. The New York Times (Feb 2022 ). Alan Rappeport and Kashmir Hill. 2022. I.R.S. to End Use of Facial Recognition for Identity Verification. The New York Times (Feb 2022)."},{"key":"e_1_3_2_2_52_1","unstructured":"Kate A Ratliff Nicole Lofaro Jennifer L Howell Morgan A Conway Calvin K Lai B O'Shea CT Smith C Jiang L Redford G Pogge etal 2020. Documenting bias from 2007--2015: Pervasiveness and correlates of implicit attitudes and stereotypes II. Unpublished Manuscript (2020).  Kate A Ratliff Nicole Lofaro Jennifer L Howell Morgan A Conway Calvin K Lai B O'Shea CT Smith C Jiang L Redford G Pogge et al. 2020. Documenting bias from 2007--2015: Pervasiveness and correlates of implicit attitudes and stereotypes II. Unpublished Manuscript (2020)."},{"key":"e_1_3_2_2_53_1","volume-title":"Racial influence on automated perceptions of emotions. Available at SSRN 3281765","author":"Rhue Lauren","year":"2018","unstructured":"Lauren Rhue . 2018. Racial influence on automated perceptions of emotions. Available at SSRN 3281765 ( 2018 ). Lauren Rhue. 2018. Racial influence on automated perceptions of emotions. Available at SSRN 3281765 (2018)."},{"key":"e_1_3_2_2_54_1","doi-asserted-by":"publisher","DOI":"10.1145\/365628.365657"},{"key":"e_1_3_2_2_55_1","unstructured":"Edward Said. 2014. Orientalism. Routledge.  Edward Said. 2014. Orientalism. Routledge."},{"key":"e_1_3_2_2_56_1","unstructured":"Christoph Schuhmann. 2021. LAION-400-Million Open Dataset. https:\/\/laion.ai\/laion-400-open-dataset\/  Christoph Schuhmann. 2021. LAION-400-Million Open Dataset. https:\/\/laion.ai\/laion-400-open-dataset\/"},{"key":"e_1_3_2_2_57_1","volume-title":"Laion-400m: Open dataset of clip-filtered 400 million image-text pairs. arXiv preprint arXiv:2111.02114","author":"Schuhmann Christoph","year":"2021","unstructured":"Christoph Schuhmann , Richard Vencu , Romain Beaumont , Robert Kaczmarczyk , Clayton Mullis , Aarush Katta , Theo Coombes , Jenia Jitsev , and Aran Komatsuzaki . 2021. Laion-400m: Open dataset of clip-filtered 400 million image-text pairs. arXiv preprint arXiv:2111.02114 ( 2021 ). Christoph Schuhmann, Richard Vencu, Romain Beaumont, Robert Kaczmarczyk, Clayton Mullis, Aarush Katta, Theo Coombes, Jenia Jitsev, and Aran Komatsuzaki. 2021. Laion-400m: Open dataset of clip-filtered 400 million image-text pairs. arXiv preprint arXiv:2111.02114 (2021)."},{"key":"e_1_3_2_2_58_1","volume-title":"Rev.","author":"Schuman Howard","year":"1997","unstructured":"Howard Schuman , Charlotte Steeh , Lawrence Bobo , and Maria Krysan . 1997. Racial attitudes in America: Trends and interpretations , Rev. ( 1997 ). Howard Schuman, Charlotte Steeh, Lawrence Bobo, and Maria Krysan. 1997. Racial attitudes in America: Trends and interpretations, Rev. (1997)."},{"key":"e_1_3_2_2_59_1","volume-title":"IRS announces transition away from use of third-party verification involving facial recognition. IRS Newsroom (Feb","author":"Internal Revenue Service U.S.","year":"2022","unstructured":"U.S. Internal Revenue Service . 2022. IRS announces transition away from use of third-party verification involving facial recognition. IRS Newsroom (Feb 2022 ). U.S. Internal Revenue Service. 2022. IRS announces transition away from use of third-party verification involving facial recognition. IRS Newsroom (Feb 2022)."},{"key":"e_1_3_2_2_60_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/P18-1238"},{"key":"e_1_3_2_2_61_1","unstructured":"Richard Socher Milind Ganjoo Christopher D Manning and Andrew Ng. 2013. Zero-Shot Learning Through Cross-Modal Transfer. In Advances in Neural Information Processing Systems. 935--943.  Richard Socher Milind Ganjoo Christopher D Manning and Andrew Ng. 2013. Zero-Shot Learning Through Cross-Modal Transfer. In Advances in Neural Information Processing Systems. 935--943."},{"key":"e_1_3_2_2_62_1","doi-asserted-by":"publisher","DOI":"10.1145\/3442188.3445932"},{"key":"e_1_3_2_2_63_1","doi-asserted-by":"publisher","DOI":"10.1145\/2812802"},{"key":"e_1_3_2_2_64_1","volume-title":"Contrastive Representation Distillation. In International Conference on Learning Representations.","author":"Tian Yonglong","year":"2019","unstructured":"Yonglong Tian , Dilip Krishnan , and Phillip Isola . 2019 . Contrastive Representation Distillation. In International Conference on Learning Representations. Yonglong Tian, Dilip Krishnan, and Phillip Isola. 2019. Contrastive Representation Distillation. In International Conference on Learning Representations."},{"key":"e_1_3_2_2_65_1","unstructured":"Saurabh Tiwary. 2021. Turing Bletchley: A Universal Image Language Representation model by Microsoft. https:\/\/www.microsoft.com\/en-us\/research\/blog\/turing-bletchley-a-universal-image-language-representation-model-by-microsoft\/  Saurabh Tiwary. 2021. Turing Bletchley: A Universal Image Language Representation model by Microsoft. https:\/\/www.microsoft.com\/en-us\/research\/blog\/turing-bletchley-a-universal-image-language-representation-model-by-microsoft\/"},{"key":"e_1_3_2_2_66_1","volume-title":"ValNorm Quantifies Semantics to Reveal Consistent Valence Biases Across Languages and Over Centuries. Empirical Methods in Natural Language Processing (EMNLP)","author":"Toney-Wails Autumn","year":"2021","unstructured":"Autumn Toney-Wails and Aylin Caliskan . 2021. ValNorm Quantifies Semantics to Reveal Consistent Valence Biases Across Languages and Over Centuries. Empirical Methods in Natural Language Processing (EMNLP) ( 2021 ). Autumn Toney-Wails and Aylin Caliskan. 2021. ValNorm Quantifies Semantics to Reveal Consistent Valence Biases Across Languages and Over Centuries. Empirical Methods in Natural Language Processing (EMNLP) (2021)."},{"key":"e_1_3_2_2_67_1","unstructured":"Ashish Vaswani Noam Shazeer Niki Parmar Jakob Uszkoreit Llion Jones Aidan N Gomez \u0141ukasz Kaiser and Illia Polosukhin. 2017. Attention is all you need. In Advances in neural information processing systems. 5998--6008.  Ashish Vaswani Noam Shazeer Niki Parmar Jakob Uszkoreit Llion Jones Aidan N Gomez \u0141ukasz Kaiser and Illia Polosukhin. 2017. Attention is all you need. In Advances in neural information processing systems. 5998--6008."},{"key":"e_1_3_2_2_68_1","volume-title":"Diachronic Analysis of German Parliamentary Proceedings: Ideological Shifts through the Lens of Political Biases. arXiv preprint arXiv:2108","author":"Walter Tobias","year":"2021","unstructured":"Tobias Walter , Celina Kirschner , Steffen Eger , Goran Glavavs , Anne Lauscher , and Simone Paolo Ponzetto . 2021 . Diachronic Analysis of German Parliamentary Proceedings: Ideological Shifts through the Lens of Political Biases. arXiv preprint arXiv:2108 .06295 (2021). Tobias Walter, Celina Kirschner, Steffen Eger, Goran Glavavs, Anne Lauscher, and Simone Paolo Ponzetto. 2021. Diachronic Analysis of German Parliamentary Proceedings: Ideological Shifts through the Lens of Political Biases. arXiv preprint arXiv:2108.06295 (2021)."},{"key":"e_1_3_2_2_69_1","volume-title":"Are Gender-Neutral Queries Really Gender-Neutral? Mitigating Gender Bias in Image Search. arXiv preprint arXiv:2109.05433","author":"Wang Jialu","year":"2021","unstructured":"Jialu Wang , Yang Liu , and Xin Eric Wang . 2021. Are Gender-Neutral Queries Really Gender-Neutral? Mitigating Gender Bias in Image Search. arXiv preprint arXiv:2109.05433 ( 2021 ). Jialu Wang, Yang Liu, and Xin Eric Wang. 2021. Are Gender-Neutral Queries Really Gender-Neutral? Mitigating Gender Bias in Image Search. arXiv preprint arXiv:2109.05433 (2021)."},{"key":"e_1_3_2_2_70_1","volume-title":"Superordinate identities and intergroup conflict: The ingroup projection model. European review of social psychology","author":"Wenzel Michael","year":"2008","unstructured":"Michael Wenzel , Am\u00e9lie Mummendey , and Sven Waldzus . 2008. Superordinate identities and intergroup conflict: The ingroup projection model. European review of social psychology , Vol. 18 , 1 ( 2008 ), 331--372. Michael Wenzel, Am\u00e9lie Mummendey, and Sven Waldzus. 2008. Superordinate identities and intergroup conflict: The ingroup projection model. European review of social psychology, Vol. 18, 1 (2008), 331--372."},{"key":"e_1_3_2_2_71_1","volume-title":"Predictive inequity in object detection. arXiv preprint arXiv:1902.11097","author":"Wilson Benjamin","year":"2019","unstructured":"Benjamin Wilson , Judy Hoffman , and Jamie Morgenstern . 2019. Predictive inequity in object detection. arXiv preprint arXiv:1902.11097 ( 2019 ). Benjamin Wilson, Judy Hoffman, and Jamie Morgenstern. 2019. Predictive inequity in object detection. arXiv preprint arXiv:1902.11097 (2019)."},{"key":"e_1_3_2_2_72_1","volume-title":"Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing: System Demonstrations. Association for Computational Linguistics, Online, 38--45","author":"Wolf Thomas","year":"2020","unstructured":"Thomas Wolf , Lysandre Debut , Victor Sanh , Julien Chaumond , Clement Delangue , Anthony Moi , Pierric Cistac , Tim Rault , R\u00e9mi Louf , Morgan Funtowicz , Joe Davison , Sam Shleifer , Patrick von Platen , Clara Ma , Yacine Jernite , Julien Plu , Canwen Xu , Teven Le Scao , Sylvain Gugger , Mariama Drame , Quentin Lhoest , and Alexander M. Rush . 2020. Transformers: State-of-the-Art Natural Language Processing . In Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing: System Demonstrations. Association for Computational Linguistics, Online, 38--45 . https:\/\/www.aclweb.org\/anthology\/ 2020 .emnlp-demos.6 Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, R\u00e9mi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander M. Rush. 2020. Transformers: State-of-the-Art Natural Language Processing. In Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing: System Demonstrations. Association for Computational Linguistics, Online, 38--45. https:\/\/www.aclweb.org\/anthology\/2020.emnlp-demos.6"},{"key":"e_1_3_2_2_73_1","volume-title":"Evidence for Hypodescent in Visual Semantic AI. ACM Conference on Fairness, Accountability, and Transparency","author":"Wolfe Robert","year":"2022","unstructured":"Robert Wolfe , Mahzarin Banaji , and Aylin Caliskan . 2022 . Evidence for Hypodescent in Visual Semantic AI. ACM Conference on Fairness, Accountability, and Transparency (2022). Robert Wolfe, Mahzarin Banaji, and Aylin Caliskan. 2022. Evidence for Hypodescent in Visual Semantic AI. ACM Conference on Fairness, Accountability, and Transparency (2022)."},{"key":"e_1_3_2_2_74_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2021.emnlp-main.41"},{"key":"e_1_3_2_2_75_1","volume-title":"2022 a. Contrastive Visual Semantic Pretraining Magnifies the Semantics of Natural Language Representations","author":"Wolfe Robert","year":"2022","unstructured":"Robert Wolfe and Aylin Caliskan . 2022 a. Contrastive Visual Semantic Pretraining Magnifies the Semantics of Natural Language Representations . Association for Computational Linguistics ( 2022 ). Robert Wolfe and Aylin Caliskan. 2022 a. Contrastive Visual Semantic Pretraining Magnifies the Semantics of Natural Language Representations. Association for Computational Linguistics (2022)."},{"key":"e_1_3_2_2_76_1","volume-title":"Markedness in Visual Semantic AI. ACM Conference on Fairness, Accountability, and Transparency","author":"Wolfe Robert","year":"2022","unstructured":"Robert Wolfe and Aylin Caliskan . 2022 b . Markedness in Visual Semantic AI. ACM Conference on Fairness, Accountability, and Transparency (2022). Robert Wolfe and Aylin Caliskan. 2022 b. Markedness in Visual Semantic AI. ACM Conference on Fairness, Accountability, and Transparency (2022)."},{"key":"e_1_3_2_2_77_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v36i10.21400"},{"key":"e_1_3_2_2_78_1","doi-asserted-by":"publisher","DOI":"10.1177\/0146167210380928"},{"key":"e_1_3_2_2_79_1","volume-title":"Contrastive learning of medical visual representations from paired images and text. arXiv preprint arXiv:2010.00747","author":"Zhang Yuhao","year":"2020","unstructured":"Yuhao Zhang , Hang Jiang , Yasuhide Miura , Christopher D Manning , and Curtis P Langlotz . 2020. Contrastive learning of medical visual representations from paired images and text. arXiv preprint arXiv:2010.00747 ( 2020 ). Yuhao Zhang, Hang Jiang, Yasuhide Miura, Christopher D Manning, and Curtis P Langlotz. 2020. Contrastive learning of medical visual representations from paired images and text. arXiv preprint arXiv:2010.00747 (2020)."},{"key":"e_1_3_2_2_80_1","doi-asserted-by":"publisher","DOI":"10.1037\/pspa0000080"}],"event":{"name":"AIES '22: AAAI\/ACM Conference on AI, Ethics, and Society","location":"Oxford United Kingdom","acronym":"AIES '22","sponsor":["SIGAI ACM Special Interest Group on Artificial Intelligence","AAAI"]},"container-title":["Proceedings of the 2022 AAAI\/ACM Conference on AI, Ethics, and Society"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3514094.3534136","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3514094.3534136","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T19:02:36Z","timestamp":1750186956000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3514094.3534136"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,7,26]]},"references-count":80,"alternative-id":["10.1145\/3514094.3534136","10.1145\/3514094"],"URL":"https:\/\/doi.org\/10.1145\/3514094.3534136","relation":{},"subject":[],"published":{"date-parts":[[2022,7,26]]},"assertion":[{"value":"2022-07-27","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}