{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,14]],"date-time":"2026-04-14T22:13:13Z","timestamp":1776204793031,"version":"3.50.1"},"publisher-location":"New York, NY, USA","reference-count":40,"publisher":"ACM","license":[{"start":{"date-parts":[[2024,6,3]],"date-time":"2024-06-03T00:00:00Z","timestamp":1717372800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2024,6,3]]},"DOI":"10.1145\/3630106.3659039","type":"proceedings-article","created":{"date-parts":[[2024,6,5]],"date-time":"2024-06-05T13:14:21Z","timestamp":1717593261000},"page":"2302-2321","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":1,"title":["Embracing Diversity: Interpretable Zero-shot Classification Beyond One Vector Per Class"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0009-0006-2842-5352","authenticated-orcid":false,"given":"Mazda","family":"Moayeri","sequence":"first","affiliation":[{"name":"University of Maryland, United States of America"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-0536-7904","authenticated-orcid":false,"given":"Michael","family":"Rabbat","sequence":"additional","affiliation":[{"name":"Meta AI, Canada"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-4972-9620","authenticated-orcid":false,"given":"Mark","family":"Ibrahim","sequence":"additional","affiliation":[{"name":"Meta AI, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0001-9078-8836","authenticated-orcid":false,"given":"Diane","family":"Bouchacourt","sequence":"additional","affiliation":[{"name":"Meta AI, Canada"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2024,6,5]]},"reference":[{"key":"e_1_3_2_1_1_1","volume-title":"More Context","author":"An Bang","unstructured":"Bang An, Sicheng Zhu, Michael-Andrei Panaitescu-Liess, Chaithanya\u00a0Kumar Mummadi, and Furong Huang. 2023. More Context, Less Distraction: Visual Classification by Inferring and Conditioning on Contextual Attributes. arxiv:2308.01313\u00a0[cs.CV]"},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-10599-4_29"},{"key":"e_1_3_2_1_3_1","volume-title":"Proceedings of the 1st Conference on Fairness, Accountability and Transparency(Proceedings of Machine Learning Research, Vol.\u00a081)","author":"Buolamwini Joy","year":"2018","unstructured":"Joy Buolamwini and Timnit Gebru. 2018. Gender Shades: Intersectional Accuracy Disparities in Commercial Gender Classification. In Proceedings of the 1st Conference on Fairness, Accountability and Transparency(Proceedings of Machine Learning Research, Vol.\u00a081), Sorelle\u00a0A. Friedler and Christo Wilson (Eds.). PMLR, 77\u201391. https:\/\/proceedings.mlr.press\/v81\/buolamwini18a.html"},{"key":"e_1_3_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.00951"},{"key":"e_1_3_2_1_5_1","volume-title":"Vicuna: An Open-Source Chatbot Impressing GPT-4 with 90%* ChatGPT Quality. https:\/\/lmsys.org\/blog\/2023-03-30-vicuna\/","author":"Chiang Wei-Lin","year":"2023","unstructured":"Wei-Lin Chiang, Zhuohan Li, Zi Lin, Ying Sheng, Zhanghao Wu, Hao Zhang, Lianmin Zheng, Siyuan Zhuang, Yonghao Zhuang, Joseph\u00a0E. Gonzalez, Ion Stoica, and Eric\u00a0P. Xing. 2023. Vicuna: An Open-Source Chatbot Impressing GPT-4 with 90%* ChatGPT Quality. https:\/\/lmsys.org\/blog\/2023-03-30-vicuna\/"},{"key":"e_1_3_2_1_6_1","volume-title":"Debiasing Vision-Language Models via Biased Prompts. arXiv preprint 2302.00070","author":"Chuang Ching-Yao","year":"2023","unstructured":"Ching-Yao Chuang, Jampani Varun, Yuanzhen Li, Antonio Torralba, and Stefanie Jegelka. 2023. Debiasing Vision-Language Models via Biased Prompts. arXiv preprint 2302.00070 (2023)."},{"key":"e_1_3_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2009.5206848"},{"key":"e_1_3_2_1_8_1","volume-title":"Cees G.\u00a0M. Snoek, Georgios Tzimiropoulos, and Brais Martinez.","author":"Derakhshani Mohammad\u00a0Mahdi","year":"2023","unstructured":"Mohammad\u00a0Mahdi Derakhshani, Enrique Sanchez, Adrian Bulat, Victor Guilherme\u00a0Turrisi da Costa, Cees G.\u00a0M. Snoek, Georgios Tzimiropoulos, and Brais Martinez. 2023. Bayesian Prompt Learning for Image-Language Model Generalization. arxiv:2210.02390\u00a0[cs.CV]"},{"key":"e_1_3_2_1_9_1","unstructured":"Zhili Feng Anna Bair and J.\u00a0Zico Kolter. 2023. Leveraging Multiple Descriptive Features for Robust Few-shot Image Learning. arxiv:2307.04317\u00a0[cs.CV]"},{"key":"e_1_3_2_1_10_1","volume-title":"The Many Faces of Robustness: A Critical Analysis of Out-of-Distribution Generalization. ICCV","author":"Hendrycks Dan","year":"2021","unstructured":"Dan Hendrycks, Steven Basart, Norman Mu, Saurav Kadavath, Frank Wang, Evan Dorundo, Rahul Desai, Tyler Zhu, Samyak Parajuli, Mike Guo, Dawn Song, Jacob Steinhardt, and Justin Gilmer. 2021. The Many Faces of Robustness: A Critical Analysis of Out-of-Distribution Generalization. ICCV (2021)."},{"key":"e_1_3_2_1_11_1","volume-title":"Natural Adversarial Examples. CVPR","author":"Hendrycks Dan","year":"2021","unstructured":"Dan Hendrycks, Kevin Zhao, Steven Basart, Jacob Steinhardt, and Dawn Song. 2021. Natural Adversarial Examples. CVPR (2021)."},{"key":"e_1_3_2_1_12_1","volume-title":"Unsupervised prompt learning for vision-language models. arXiv preprint arXiv:2204.03649","author":"Huang Tony","year":"2022","unstructured":"Tony Huang, Jack Chu, and Fangyun Wei. 2022. Unsupervised prompt learning for vision-language models. arXiv preprint arXiv:2204.03649 (2022)."},{"key":"e_1_3_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2015.7298744"},{"key":"e_1_3_2_1_14_1","unstructured":"Younghyun Kim Sangwoo Mo Minkyu Kim Kyungmin Lee Jaeho Lee and Jinwoo Shin. 2023. Bias-to-Text: Debiasing Unknown Visual Biases through Language Interpretation. arxiv:2301.11104\u00a0[cs.LG]"},{"key":"e_1_3_2_1_15_1","unstructured":"Junnan Li Dongxu Li Silvio Savarese and Steven Hoi. 2023. BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models. arxiv:2301.12597\u00a0[cs.CV]"},{"key":"e_1_3_2_1_16_1","unstructured":"Subhransu Maji Esa Rahtu Juho Kannala Matthew Blaschko and Andrea Vedaldi. 2013. Fine-Grained Visual Classification of Aircraft. arxiv:1306.5151\u00a0[cs.CV]"},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.48550\/arXiv.2306.01669"},{"key":"e_1_3_2_1_18_1","volume-title":"Visual Classification via Description from Large Language Models. ICLR","author":"Menon Sachit","year":"2023","unstructured":"Sachit Menon and Carl Vondrick. 2023. Visual Classification via Description from Large Language Models. ICLR (2023)."},{"key":"e_1_3_2_1_19_1","volume-title":"LaFTer: Label-Free Tuning of Zero-shot Classifier using Language and Unlabeled Image Collections. arXiv:2305.18287 (May","author":"Mirza Jehanzeb","year":"2023","unstructured":"M.\u00a0Jehanzeb Mirza, Leonid Karlinsky, Wei Lin, Mateusz Kozinski, Horst Possegger, Rogerio Feris, and Horst Bischof. 2023. LaFTer: Label-Free Tuning of Zero-shot Classifier using Language and Unlabeled Image Collections. arXiv:2305.18287 (May 2023). http:\/\/arxiv.org\/abs\/2305.18287 arXiv:2305.18287 [cs]."},{"key":"e_1_3_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICVGIP.2008.47"},{"key":"e_1_3_2_1_21_1","volume-title":"CHiLS: Zero-Shot Image Classification with Hierarchical Label Sets. In International Conference on Machine Learning (ICML).","author":"Novack Zachary","year":"2023","unstructured":"Zachary Novack, Julian McAuley, Zachary Lipton, and Saurabh Garg. 2023. CHiLS: Zero-Shot Image Classification with Hierarchical Label Sets. In International Conference on Machine Learning (ICML)."},{"key":"e_1_3_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2012.6248092"},{"key":"e_1_3_2_1_23_1","doi-asserted-by":"crossref","unstructured":"Sarah\u00a0M Pratt Rosanne Liu and Ali Farhadi. 2023. What does a platypus look like? Generating customized prompts for zero-shot image classification. https:\/\/openreview.net\/forum?id=3ly9cG9Ql9h","DOI":"10.1109\/ICCV51070.2023.01438"},{"key":"e_1_3_2_1_24_1","unstructured":"Alec Radford Jong\u00a0Wook Kim Chris Hallacy Aditya Ramesh Gabriel Goh Sandhini Agarwal Girish Sastry Amanda Askell Pamela Mishkin Jack Clark Gretchen Krueger and Ilya Sutskever. 2021. Learning Transferable Visual Models From Natural Language Supervision. arxiv:2103.00020\u00a0[cs.CV]"},{"key":"e_1_3_2_1_25_1","doi-asserted-by":"crossref","unstructured":"Vikram\u00a0V. Ramaswamy Sunnie S.\u00a0Y. Kim Ruth Fong and Olga Russakovsky. 2023. Overlooked factors in concept-based explanations: Dataset choice concept learnability and human capability. arxiv:2207.09615\u00a0[cs.CV]","DOI":"10.1109\/CVPR52729.2023.01052"},{"key":"e_1_3_2_1_26_1","volume-title":"Deepti Ghadiyaram, and Olga Russakovsky.","author":"Ramaswamy V.","year":"2023","unstructured":"Vikram\u00a0V. Ramaswamy, Sing\u00a0Yu Lin, Dora Zhao, Aaron\u00a0B. Adcock, Laurens van\u00a0der Maaten, Deepti Ghadiyaram, and Olga Russakovsky. 2023. Beyond web-scraping: Crowd-sourcing a geodiverse dataset. In arXiv preprint."},{"key":"e_1_3_2_1_27_1","unstructured":"Benjamin Recht Rebecca Roelofs Ludwig Schmidt and Vaishaal Shankar. 2019. Do ImageNet Classifiers Generalize to ImageNet?arxiv:1902.10811\u00a0[cs.CV]"},{"key":"e_1_3_2_1_28_1","unstructured":"Megan Richards Polina Kirichenko Diane Bouchacourt and Mark Ibrahim. 2023. Does Progress On Object Recognition Benchmarks Improve Real-World Generalization?arxiv:2307.13136\u00a0[cs.CV]"},{"key":"e_1_3_2_1_29_1","unstructured":"William A\u00a0Gaviria Rojas Sudnya Diamos Keertan\u00a0Ranjan Kini David Kanter Vijay\u00a0Janapa Reddi and Cody Coleman. 2022. The Dollar Street Dataset: Images Representing the Geographic and Socioeconomic Diversity of the World. In Thirty-sixth Conference on Neural Information Processing Systems Datasets and Benchmarks Track. https:\/\/openreview.net\/forum?id=qnfYsave0U4"},{"key":"e_1_3_2_1_30_1","doi-asserted-by":"crossref","unstructured":"Karsten Roth Jae\u00a0Myung Kim A.\u00a0Sophia Koepke Oriol Vinyals Cordelia Schmid and Zeynep Akata. 2023. Waffling around for Performance: Visual Classification with Random Words and Broad Concepts. arxiv:2306.07282\u00a0[cs.CV]","DOI":"10.1109\/ICCV51070.2023.01443"},{"key":"e_1_3_2_1_31_1","volume-title":"BREEDS: Benchmarks for Subpopulation Shift. arXiv: Computer Vision and Pattern Recognition","author":"Santurkar Shibani","year":"2020","unstructured":"Shibani Santurkar, Dimitris Tsipras, and Aleksander Madry. 2020. BREEDS: Benchmarks for Subpopulation Shift. arXiv: Computer Vision and Pattern Recognition (2020). https:\/\/api.semanticscholar.org\/CorpusID:221095529"},{"key":"e_1_3_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52729.2023.00659"},{"key":"e_1_3_2_1_33_1","volume-title":"International Conference on Machine Learning, Vol.\u00a0139","author":"Touvron Hugo","year":"2021","unstructured":"Hugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa, Alexandre Sablayrolles, and Herve Jegou. 2021. Training data-efficient image transformers I& distillation through attention. In International Conference on Machine Learning, Vol.\u00a0139. 10347\u201310357."},{"key":"e_1_3_2_1_34_1","unstructured":"Haohan Wang Songwei Ge Zachary Lipton and Eric\u00a0P Xing. 2019. Learning Robust Global Representations by Penalizing Local Predictive Power. In Advances in Neural Information Processing Systems. 10506\u201310518."},{"key":"e_1_3_2_1_35_1","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR). 19187\u201319197","author":"Yang Yue","year":"2023","unstructured":"Yue Yang, Artemis Panagopoulou, Shenghao Zhou, Daniel Jin, Chris Callison-Burch, and Mark Yatskar. 2023. Language in a Bottle: Language Model Guided Concept Bottlenecks for Interpretable Image Classification. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR). 19187\u201319197."},{"key":"e_1_3_2_1_36_1","volume-title":"CoCa: Contrastive Captioners are Image-Text Foundation Models. Trans. Mach. Learn. Res. 2022","author":"Yu Jiahui","year":"2022","unstructured":"Jiahui Yu, Zirui Wang, Vijay Vasudevan, Legg Yeung, Mojtaba Seyedhosseini, and Yonghui Wu. 2022. CoCa: Contrastive Captioners are Image-Text Foundation Models. Trans. Mach. Learn. Res. 2022 (2022). https:\/\/api.semanticscholar.org\/CorpusID:248512473"},{"key":"e_1_3_2_1_37_1","volume-title":"The Eleventh International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=D-zfUK7BR6c","author":"Zhang Yuhui","year":"2023","unstructured":"Yuhui Zhang, Jeff\u00a0Z. HaoChen, Shih-Cheng Huang, Kuan-Chieh Wang, James Zou, and Serena Yeung. 2023. Diagnosing and Rectifying Vision Models using Language. In The Eleventh International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=D-zfUK7BR6c"},{"key":"e_1_3_2_1_38_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.01631"},{"key":"e_1_3_2_1_39_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-022-01653-1"},{"key":"e_1_3_2_1_40_1","volume-title":"Prompt-aligned gradient for prompt tuning. arXiv preprint arXiv:2205.14865","author":"Zhu Beier","year":"2022","unstructured":"Beier Zhu, Yulei Niu, Yucheng Han, Yue Wu, and Hanwang Zhang. 2022. Prompt-aligned gradient for prompt tuning. arXiv preprint arXiv:2205.14865 (2022)."}],"event":{"name":"FAccT '24: The 2024 ACM Conference on Fairness, Accountability, and Transparency","location":"Rio de Janeiro Brazil","acronym":"FAccT '24"},"container-title":["The 2024 ACM Conference on Fairness, Accountability, and Transparency"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3630106.3659039","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3630106.3659039","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T23:57:07Z","timestamp":1750291027000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3630106.3659039"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,6,3]]},"references-count":40,"alternative-id":["10.1145\/3630106.3659039","10.1145\/3630106"],"URL":"https:\/\/doi.org\/10.1145\/3630106.3659039","relation":{},"subject":[],"published":{"date-parts":[[2024,6,3]]},"assertion":[{"value":"2024-06-05","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}