{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,8,29]],"date-time":"2025-08-29T00:05:14Z","timestamp":1756425914670,"version":"3.44.0"},"publisher-location":"New York, NY, USA","reference-count":27,"publisher":"ACM","funder":[{"name":"Staatskanzlei des Saarlandes","award":["PABeLA"],"award-info":[{"award-number":["PABeLA"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2025,8,31]]},"DOI":"10.1145\/3743049.3748541","type":"proceedings-article","created":{"date-parts":[[2025,8,28]],"date-time":"2025-08-28T14:03:01Z","timestamp":1756389781000},"page":"653-658","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["Identification of Important LLM Test Criteria for State Media Authorities"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0009-0007-4788-6848","authenticated-orcid":false,"given":"Stefan","family":"Schaffer","sequence":"first","affiliation":[{"name":"German Research Center for Artificial Intelligence (DFKI), Berlin, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-4288-4724","authenticated-orcid":false,"given":"Niko","family":"Kleer","sequence":"additional","affiliation":[{"name":"German Research Center for Artificial Intelligence (DFKI), Saarland Informatics Campus, Saarbr\u00fccken, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-5117-6928","authenticated-orcid":false,"given":"Pascal","family":"Lessel","sequence":"additional","affiliation":[{"name":"German Research Center for Artificial Intelligence (DFKI), Saarland Informatics Campus, Saarbr\u00fccken, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-6755-5287","authenticated-orcid":false,"given":"Michael","family":"Feld","sequence":"additional","affiliation":[{"name":"German Research Center for Artificial Intelligence (DFKI), Saarland Informatics Campus, Saarbr\u00fccken, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0005-5241-3100","authenticated-orcid":false,"given":"Julien Armin","family":"Bauer","sequence":"additional","affiliation":[{"name":"Landesmedienanstalt Saarland (LMS), Saarbr\u00fccken, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0000-4271-8617","authenticated-orcid":false,"given":"Ina","family":"Goedert","sequence":"additional","affiliation":[{"name":"Landesmedienanstalt Saarland (LMS), Saarbr\u00fccken, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2025,8,30]]},"reference":[{"key":"e_1_3_3_2_2_2","unstructured":"Alok Abhishek Lisa Erickson and Tushar Bandopadhyay. 2025. BEATS: Bias Evaluation and Assessment Test Suite for Large Language Models. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2503.24310 (2025)."},{"key":"e_1_3_3_2_3_2","unstructured":"Rishi Bommasani Drew\u00a0A Hudson Ehsan Adeli Russ Altman Simran Arora Sydney von Arx Michael\u00a0S Bernstein Jeannette Bohg Antoine Bosselut Emma Brunskill et\u00a0al. 2021. On the opportunities and risks of foundation models. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2108.07258 (2021)."},{"key":"e_1_3_3_2_4_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2024.findings-acl.304"},{"key":"e_1_3_3_2_5_2","doi-asserted-by":"publisher","DOI":"10.1109\/SaTML64287.2025.00010"},{"key":"e_1_3_3_2_6_2","unstructured":"Veronica Chatrath Oluwanifemi Bamgbose and Shaina Raza. 2023. She had Cobalt Blue Eyes: Prompt Testing to Create Aligned and Sustainable Language Models. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2310.18333 (2023)."},{"key":"e_1_3_3_2_7_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2024.naacl-long.118"},{"key":"e_1_3_3_2_8_2","doi-asserted-by":"publisher","DOI":"10.1145\/3630106.3658939"},{"key":"e_1_3_3_2_9_2","doi-asserted-by":"crossref","unstructured":"Isabel\u00a0O. Gallegos Ryan\u00a0A. Rossi Joe Barrow Md\u00a0Mehrab Tanjim Sungchul Kim Franck Dernoncourt Tong Yu Ruiyi Zhang and Nesreen\u00a0K. Ahmed. 2024. Bias and Fairness in Large Language Models: A Survey. Computational Linguistics 50 3 (2024) 1097\u20131179.","DOI":"10.1162\/coli_a_00524"},{"key":"e_1_3_3_2_10_2","unstructured":"Deep Ganguli Liane Lovitt Jackson Kernion Amanda Askell Yuntao Bai Saurav Kadavath Ben Mann Ethan Perez Nicholas Schiefer Kamal Ndousse et\u00a0al. 2022. Red Teaming Language Models to Reduce Harms: Methods Scaling Behaviors and Lessons Learned. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2209.07858 (2022)."},{"key":"e_1_3_3_2_11_2","doi-asserted-by":"crossref","unstructured":"Don Gotterbarn Amy Bruckman Catherine Flick Keith Miller and Marty\u00a0J. Wolf. 2017. ACM Code of Ethics: a Guide for Positive Action. 121\u2013128\u00a0pages.","DOI":"10.1145\/3173016"},{"key":"e_1_3_3_2_12_2","doi-asserted-by":"crossref","unstructured":"Richard Heersmink Barend de Rooij Mar\u00eda\u00a0Jimena Clavel\u00a0V\u00e1zquez and Matteo Colombo. 2024. A phenomenology and epistemology of large language models: Transparency trust and trustworthiness. Ethics and Information Technology 26 3 (2024) 41.","DOI":"10.1007\/s10676-024-09777-3"},{"key":"e_1_3_3_2_13_2","volume-title":"Ethically Aligned Design: A Vision for Prioritizing Human Well-being with Autonomous and Intelligent Systems","author":"Initiative The IEEE\u00a0Global","year":"2019","unstructured":"The IEEE\u00a0Global Initiative. 2019. Ethically Aligned Design: A Vision for Prioritizing Human Well-being with Autonomous and Intelligent Systems. https:\/\/standards.ieee.org\/wp-content\/uploads\/import\/documents\/other\/ead_v2.pdf"},{"key":"e_1_3_3_2_14_2","unstructured":"Anna Kruspe. 2024. Towards Detecting Unanticipated Bias in Large Language Models. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2404.02650 (2024)."},{"key":"e_1_3_3_2_15_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2024.findings-acl.198"},{"key":"e_1_3_3_2_16_2","unstructured":"Shi Lin Rongchang Li Xun Wang Changting Lin Wenpeng Xing and Meng Han. 2024. Figure it Out: Analyzing-based Jailbreak Attack on Large Language Models. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2407.16205 (2024)."},{"key":"e_1_3_3_2_17_2","volume-title":"The Twelfth International Conference on Learning Representations","author":"Liu Xiaogeng","year":"2024","unstructured":"Xiaogeng Liu, Nan Xu, Muhao Chen, and Chaowei Xiao. 2024. AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models. In The Twelfth International Conference on Learning Representations."},{"key":"e_1_3_3_2_18_2","first-page":"35181","volume-title":"Proceedings of the 41st International Conference on Machine Learning","author":"Mazeika Mantas","year":"2024","unstructured":"Mantas Mazeika, Long Phan, Xuwang Yin, Andy Zou, Zifan Wang, Norman Mu, Elham Sakhaee, Nathaniel Li, Steven Basart, Bo Li, et\u00a0al. 2024. HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal. In Proceedings of the 41st International Conference on Machine Learning. 35181\u201335224."},{"key":"e_1_3_3_2_19_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2022.emnlp-main.225"},{"key":"e_1_3_3_2_20_2","volume-title":"PromptFoo: LLM Prompt Engineering Toolkit","year":"2024","unstructured":"PromptFoo. 2024. PromptFoo: LLM Prompt Engineering Toolkit. https:\/\/www.promptfoo.dev\/"},{"key":"e_1_3_3_2_21_2","unstructured":"Xiongtao Sun Deyue Zhang Dongdong Yang Quanchen Zou and Hui Li. 2024. Multi-Turn Context Jailbreak Attack on Large Language Models From First Principles. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2408.04686 (2024)."},{"key":"e_1_3_3_2_22_2","doi-asserted-by":"publisher","DOI":"10.1145\/3658644.3670284"},{"key":"e_1_3_3_2_23_2","doi-asserted-by":"crossref","unstructured":"Zeguan Xiao Yan Yang Guanhua Chen and Yun Chen. 2024. Tastle: Distract Large Language Models for Automatic Jailbreak Attack. arXiv preprint arXiv.2403.08424 (2024).","DOI":"10.18653\/v1\/2024.emnlp-main.908"},{"key":"e_1_3_3_2_24_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2025.naacl-long.42"},{"key":"e_1_3_3_2_25_2","unstructured":"Jiahao Yu Xingwei Lin Zheng Yu and Xinyu Xing. 2023. Gptfuzzer: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2309.10253 (2023)."},{"key":"e_1_3_3_2_26_2","unstructured":"Tianrong Zhang Bochuan Cao Yuanpu Cao Lu Lin Prasenjit Mitra and Jinghui Chen. 2024. Wordgame: Efficient & Effective LLM Jailbreak via Simultaneous Obfuscation in Query and Response. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2405.14023 (2024)."},{"key":"e_1_3_3_2_27_2","unstructured":"Jiaxu Zhao Meng Fang Shirui Pan Wenpeng Yin and Mykola Pechenizkiy. 2023. Gptbias: A Comprehensive Framework for Evaluating Bias in Large Language Models. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2312.06315 (2023)."},{"key":"e_1_3_3_2_28_2","unstructured":"Weikang Zhou Xiao Wang Limao Xiong Han Xia Yingshuang Gu Mingxu Chai Fukang Zhu Caishuang Huang Shihan Dou Zhiheng Xi et\u00a0al. 2024. Easyjailbreak: A Unified Framework for Jailbreaking Large Language Models. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2403.12171 (2024)."}],"event":{"name":"MuC '25: Mensch und Computer 2025","location":"Chemnitz Germany","acronym":"MuC '25"},"container-title":["Proceedings of the Mensch und Computer 2025"],"original-title":[],"deposited":{"date-parts":[[2025,8,28]],"date-time":"2025-08-28T14:56:00Z","timestamp":1756392960000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3743049.3748541"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,8,30]]},"references-count":27,"alternative-id":["10.1145\/3743049.3748541","10.1145\/3743049"],"URL":"https:\/\/doi.org\/10.1145\/3743049.3748541","relation":{},"subject":[],"published":{"date-parts":[[2025,8,30]]},"assertion":[{"value":"2025-08-30","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}