{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,3]],"date-time":"2026-07-03T16:39:25Z","timestamp":1783096765664,"version":"3.54.6"},"publisher-location":"New York, NY, USA","reference-count":41,"publisher":"ACM","license":[{"start":{"date-parts":[[2020,8,20]],"date-time":"2020-08-20T00:00:00Z","timestamp":1597881600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"The TuringShield team of Tencent"},{"name":"AWS Machine Learning for Research Award"},{"name":"Google Faculty Research Award"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2020,8,23]]},"DOI":"10.1145\/3394486.3403071","type":"proceedings-article","created":{"date-parts":[[2020,8,20]],"date-time":"2020-08-20T23:18:56Z","timestamp":1597965536000},"page":"286-296","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":12,"title":["Adversarial Infidelity Learning for Model Interpretation"],"prefix":"10.1145","author":[{"given":"Jian","family":"Liang","sequence":"first","affiliation":[{"name":"Tencent, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Bing","family":"Bai","sequence":"additional","affiliation":[{"name":"Tencent, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yuren","family":"Cao","sequence":"additional","affiliation":[{"name":"Tencent, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Kun","family":"Bai","sequence":"additional","affiliation":[{"name":"Tencent, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Fei","family":"Wang","sequence":"additional","affiliation":[{"name":"Weill Cornell Medicine, New York, NY, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2020,8,20]]},"reference":[{"key":"e_1_3_2_2_1_1","first-page":"265","article-title":"Tensorflow: a system for large-scale machine learning","volume":"16","author":"Abadi Mart'in","year":"2016","unstructured":"Mart'in Abadi , Paul Barham , Jianmin Chen , Zhifeng Chen , Andy Davis , Jeffrey Dean , Matthieu Devin , Sanjay Ghemawat , Geoffrey Irving , Michael Isard , 2016 . Tensorflow: a system for large-scale machine learning .. In OSDI , Vol. 16. 265 -- 283 . Mart'in Abadi, Paul Barham, Jianmin Chen, Zhifeng Chen, Andy Davis, Jeffrey Dean, Matthieu Devin, Sanjay Ghemawat, Geoffrey Irving, Michael Isard, et al. 2016. Tensorflow: a system for large-scale machine learning.. In OSDI, Vol. 16. 265--283.","journal-title":"OSDI"},{"key":"e_1_3_2_2_2_1","volume-title":"Local explanation methods for deep neural networks lack sensitivity to parameter values. arXiv preprint arXiv:1810.03307","author":"Adebayo Julius","year":"2018","unstructured":"Julius Adebayo , Justin Gilmer , Ian Goodfellow , and Been Kim . 2018a. Local explanation methods for deep neural networks lack sensitivity to parameter values. arXiv preprint arXiv:1810.03307 ( 2018 ). Julius Adebayo, Justin Gilmer, Ian Goodfellow, and Been Kim. 2018a. Local explanation methods for deep neural networks lack sensitivity to parameter values. arXiv preprint arXiv:1810.03307 (2018)."},{"key":"e_1_3_2_2_3_1","unstructured":"Julius Adebayo Justin Gilmer Michael Muelly Ian Goodfellow Moritz Hardt and Been Kim. 2018b. Sanity checks for saliency maps. In Advances in Neural Information Processing Systems. 9505--9515.  Julius Adebayo Justin Gilmer Michael Muelly Ian Goodfellow Moritz Hardt and Been Kim. 2018b. Sanity checks for saliency maps. In Advances in Neural Information Processing Systems. 9505--9515."},{"key":"e_1_3_2_2_4_1","volume-title":"NIPS 2017-Workshop on Interpreting, Explaining and Visualizing Deep Learning. ETH Zurich.","author":"Ancona Marco","year":"2017","unstructured":"Marco Ancona , Enea Ceolini , Cengiz \u00d6ztireli , and Markus Gross . 2017 . A unified view of gradient-based attribution methods for deep neural networks . In NIPS 2017-Workshop on Interpreting, Explaining and Visualizing Deep Learning. ETH Zurich. Marco Ancona, Enea Ceolini, Cengiz \u00d6ztireli, and Markus Gross. 2017. A unified view of gradient-based attribution methods for deep neural networks. In NIPS 2017-Workshop on Interpreting, Explaining and Visualizing Deep Learning. ETH Zurich."},{"key":"e_1_3_2_2_5_1","volume-title":"Survey and critique of techniques for extracting rules from trained artificial neural networks. Knowledge-based systems","author":"Andrews Robert","year":"1995","unstructured":"Robert Andrews , Joachim Diederich , and Alan B Tickle . 1995. Survey and critique of techniques for extracting rules from trained artificial neural networks. Knowledge-based systems , Vol. 8 , 6 ( 1995 ), 373--389. Robert Andrews, Joachim Diederich, and Alan B Tickle. 1995. Survey and critique of techniques for extracting rules from trained artificial neural networks. Knowledge-based systems, Vol. 8, 6 (1995), 373--389."},{"key":"e_1_3_2_2_6_1","doi-asserted-by":"publisher","DOI":"10.1371\/journal.pone.0130140"},{"key":"e_1_3_2_2_7_1","volume-title":"Explaining a black-box using Deep Variational Information Bottleneck Approach. arXiv preprint arXiv:1902.06918","author":"Bang Seojin","year":"2019","unstructured":"Seojin Bang , Pengtao Xie , Wei Wu , and Eric Xing . 2019. Explaining a black-box using Deep Variational Information Bottleneck Approach. arXiv preprint arXiv:1902.06918 ( 2019 ). Seojin Bang, Pengtao Xie, Wei Wu, and Eric Xing. 2019. Explaining a black-box using Deep Variational Information Bottleneck Approach. arXiv preprint arXiv:1902.06918 (2019)."},{"key":"e_1_3_2_2_8_1","volume-title":"Proceedings of the International Conference on Automated Planning and Scheduling","volume":"29","author":"Chakraborti Tathagata","year":"2019","unstructured":"Tathagata Chakraborti , Anagha Kulkarni , Sarath Sreedharan , David E Smith , and Subbarao Kambhampati . 2019 . Explicability? legibility? predictability? transparency? privacy? security? the emerging landscape of interpretable agent behavior . In Proceedings of the International Conference on Automated Planning and Scheduling , Vol. 29 . 86--96. Tathagata Chakraborti, Anagha Kulkarni, Sarath Sreedharan, David E Smith, and Subbarao Kambhampati. 2019. Explicability? legibility? predictability? transparency? privacy? security? the emerging landscape of interpretable agent behavior. In Proceedings of the International Conference on Automated Planning and Scheduling, Vol. 29. 86--96."},{"key":"e_1_3_2_2_9_1","volume-title":"International Conference on Machine Learning. 883--892","author":"Chen Jianbo","year":"2018","unstructured":"Jianbo Chen , Le Song , Martin Wainwright , and Michael Jordan . 2018 . Learning to Explain: An Information-Theoretic Perspective on Model Interpretation . In International Conference on Machine Learning. 883--892 . Jianbo Chen, Le Song, Martin Wainwright, and Michael Jordan. 2018. Learning to Explain: An Information-Theoretic Perspective on Model Interpretation. In International Conference on Machine Learning. 883--892."},{"key":"e_1_3_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2009.5206848"},{"key":"e_1_3_2_2_11_1","volume-title":"Advances in Neural Information Processing Systems 32, H. Wallach, H. Larochelle, A. Beygelzimer, F. dtextquotesingle Alch\u00e9-Buc","author":"Dombrowski Ann-Kathrin","unstructured":"Ann-Kathrin Dombrowski , Maximillian Alber , Christopher Anders , Marcel Ackermann , Klaus-Robert M\u00fcller , and Pan Kessel . 2019. Explanations can be manipulated and geometry is to blame . In Advances in Neural Information Processing Systems 32, H. Wallach, H. Larochelle, A. Beygelzimer, F. dtextquotesingle Alch\u00e9-Buc , E. Fox, and R. Garnett (Eds.). Curran Associates, Inc. , 13567--13578. http:\/\/papers.nips.cc\/paper\/9511-explanations-can-be-manipulated-and-geometry-is-to-blame.pdf Ann-Kathrin Dombrowski, Maximillian Alber, Christopher Anders, Marcel Ackermann, Klaus-Robert M\u00fcller, and Pan Kessel. 2019. Explanations can be manipulated and geometry is to blame. In Advances in Neural Information Processing Systems 32, H. Wallach, H. Larochelle, A. Beygelzimer, F. dtextquotesingle Alch\u00e9-Buc, E. Fox, and R. Garnett (Eds.). Curran Associates, Inc., 13567--13578. http:\/\/papers.nips.cc\/paper\/9511-explanations-can-be-manipulated-and-geometry-is-to-blame.pdf"},{"key":"e_1_3_2_2_12_1","doi-asserted-by":"publisher","DOI":"10.1145\/3359786"},{"key":"e_1_3_2_2_13_1","unstructured":"Satoshi Hara Koichi Ikeno Tasuku Soma and Takanori Maehara. 2019. Feature Attribution As Feature Selection. https:\/\/openreview.net\/forum?id=H1lS8oA5YQ  Satoshi Hara Koichi Ikeno Tasuku Soma and Takanori Maehara. 2019. Feature Attribution As Feature Selection. https:\/\/openreview.net\/forum?id=H1lS8oA5YQ"},{"key":"e_1_3_2_2_14_1","volume-title":"Advances in Neural Information Processing Systems 32, H. Wallach, H. Larochelle, A. Beygelzimer, F. d'Alch\u00e9-Buc","author":"Heo Juyeon","unstructured":"Juyeon Heo , Sunghwan Joo , and Taesup Moon . 2019. Fooling Neural Network Interpretations via Adversarial Model Manipulation . In Advances in Neural Information Processing Systems 32, H. Wallach, H. Larochelle, A. Beygelzimer, F. d'Alch\u00e9-Buc , E. Fox, and R. Garnett (Eds.). Curran Associates, Inc. , 2921--2932. http:\/\/papers.nips.cc\/paper\/8558-fooling-neural-network-interpretations-via-adversarial-model-manipulation.pdf Juyeon Heo, Sunghwan Joo, and Taesup Moon. 2019. Fooling Neural Network Interpretations via Adversarial Model Manipulation. In Advances in Neural Information Processing Systems 32, H. Wallach, H. Larochelle, A. Beygelzimer, F. d'Alch\u00e9-Buc, E. Fox, and R. Garnett (Eds.). Curran Associates, Inc., 2921--2932. http:\/\/papers.nips.cc\/paper\/8558-fooling-neural-network-interpretations-via-adversarial-model-manipulation.pdf"},{"key":"e_1_3_2_2_15_1","unstructured":"Sara Hooker Dumitru Erhan Pieter-Jan Kindermans and Been Kim. 2019. A benchmark for interpretability methods in deep neural networks. In Advances in Neural Information Processing Systems. 9734--9745.  Sara Hooker Dumitru Erhan Pieter-Jan Kindermans and Been Kim. 2019. A benchmark for interpretability methods in deep neural networks. In Advances in Neural Information Processing Systems. 9734--9745."},{"key":"e_1_3_2_2_16_1","volume-title":"Mobilenets: Efficient convolutional neural networks for mobile vision applications. arXiv preprint arXiv:1704.04861","author":"Howard Andrew G","year":"2017","unstructured":"Andrew G Howard , Menglong Zhu , Bo Chen , Dmitry Kalenichenko , Weijun Wang , Tobias Weyand , Marco Andreetto , and Hartwig Adam . 2017 . Mobilenets: Efficient convolutional neural networks for mobile vision applications. arXiv preprint arXiv:1704.04861 (2017). Andrew G Howard, Menglong Zhu, Bo Chen, Dmitry Kalenichenko, Weijun Wang, Tobias Weyand, Marco Andreetto, and Hartwig Adam. 2017. Mobilenets: Efficient convolutional neural networks for mobile vision applications. arXiv preprint arXiv:1704.04861 (2017)."},{"key":"e_1_3_2_2_17_1","volume-title":"Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies","volume":"1","author":"Jain Sarthak","year":"2019","unstructured":"Sarthak Jain and Byron C Wallace . 2019 . Attention is not Explanation . In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies , Volume 1 (Long and Short Papers). 3543--3556. Sarthak Jain and Byron C Wallace. 2019. Attention is not Explanation. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers). 3543--3556."},{"key":"e_1_3_2_2_18_1","volume-title":"The relativistic discriminator: a key element missing from standard GAN. arXiv preprint arXiv:1807.00734","author":"Jolicoeur-Martineau Alexia","year":"2018","unstructured":"Alexia Jolicoeur-Martineau . 2018. The relativistic discriminator: a key element missing from standard GAN. arXiv preprint arXiv:1807.00734 ( 2018 ). Alexia Jolicoeur-Martineau. 2018. The relativistic discriminator: a key element missing from standard GAN. arXiv preprint arXiv:1807.00734 (2018)."},{"key":"e_1_3_2_2_19_1","volume-title":"Seong Tae Kim, and Nassir Navab","author":"Khakzar Ashkan","year":"2019","unstructured":"Ashkan Khakzar , Soroosh Baselizadeh , Saurabh Khanduja , Seong Tae Kim, and Nassir Navab . 2019 . Explaining Neural Networks via Perturbing Important Learned Features . arXiv preprint arXiv:1911.11081 (2019). Ashkan Khakzar, Soroosh Baselizadeh, Saurabh Khanduja, Seong Tae Kim, and Nassir Navab. 2019. Explaining Neural Networks via Perturbing Important Learned Features. arXiv preprint arXiv:1911.11081 (2019)."},{"key":"e_1_3_2_2_20_1","doi-asserted-by":"publisher","DOI":"10.1109\/5.726791"},{"key":"e_1_3_2_2_21_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.01053"},{"key":"e_1_3_2_2_22_1","doi-asserted-by":"publisher","DOI":"10.1145\/3236386.3241340"},{"key":"e_1_3_2_2_23_1","unstructured":"Scott M Lundberg and Su-In Lee. 2017. A unified approach to interpreting model predictions. In Advances in Neural Information Processing Systems. 4765--4774.  Scott M Lundberg and Su-In Lee. 2017. A unified approach to interpreting model predictions. In Advances in Neural Information Processing Systems. 4765--4774."},{"key":"e_1_3_2_2_24_1","doi-asserted-by":"publisher","DOI":"10.5555\/2002472.2002491"},{"key":"e_1_3_2_2_25_1","doi-asserted-by":"publisher","DOI":"10.1145\/3316781.3317822"},{"key":"e_1_3_2_2_26_1","unstructured":"Gregory Plumb Denali Molitor and Ameet S Talwalkar. 2018. Model agnostic supervised local explanations. In Advances in Neural Information Processing Systems. 2515--2524.  Gregory Plumb Denali Molitor and Ameet S Talwalkar. 2018. Model agnostic supervised local explanations. In Advances in Neural Information Processing Systems. 2515--2524."},{"key":"e_1_3_2_2_27_1","doi-asserted-by":"publisher","DOI":"10.1145\/2939672.2939778"},{"key":"e_1_3_2_2_28_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-24574-4_28"},{"key":"e_1_3_2_2_29_1","volume-title":"Eleventh Artificial Intelligence and Interactive Digital Entertainment Conference .","author":"Schwab Patrick","year":"2015","unstructured":"Patrick Schwab and Helmut Hlavacs . 2015 . Capturing the essence: Towards the automated generation of transparent behavior models . In Eleventh Artificial Intelligence and Interactive Digital Entertainment Conference . Patrick Schwab and Helmut Hlavacs. 2015. Capturing the essence: Towards the automated generation of transparent behavior models. In Eleventh Artificial Intelligence and Interactive Digital Entertainment Conference ."},{"key":"e_1_3_2_2_30_1","unstructured":"Patrick Schwab and Walter Karlen. 2019. CXPlain: Causal explanations for model interpretation under uncertainty. In Advances in Neural Information Processing Systems. 10220--10230.  Patrick Schwab and Walter Karlen. 2019. CXPlain: Causal explanations for model interpretation under uncertainty. In Advances in Neural Information Processing Systems. 10220--10230."},{"key":"e_1_3_2_2_31_1","doi-asserted-by":"publisher","DOI":"10.5555\/3305890.3306006"},{"key":"e_1_3_2_2_32_1","volume-title":"Deep inside convolutional networks: Visualising image classification models and saliency maps. arXiv preprint arXiv:1312.6034","author":"Simonyan Karen","year":"2013","unstructured":"Karen Simonyan , Andrea Vedaldi , and Andrew Zisserman . 2013. Deep inside convolutional networks: Visualising image classification models and saliency maps. arXiv preprint arXiv:1312.6034 ( 2013 ). Karen Simonyan, Andrea Vedaldi, and Andrew Zisserman. 2013. Deep inside convolutional networks: Visualising image classification models and saliency maps. arXiv preprint arXiv:1312.6034 (2013)."},{"key":"e_1_3_2_2_33_1","volume-title":"Smoothgrad: removing noise by adding noise. arXiv preprint arXiv:1706.03825","author":"Smilkov Daniel","year":"2017","unstructured":"Daniel Smilkov , Nikhil Thorat , Been Kim , Fernanda Vi\u00e9gas , and Martin Wattenberg . 2017. Smoothgrad: removing noise by adding noise. arXiv preprint arXiv:1706.03825 ( 2017 ). Daniel Smilkov, Nikhil Thorat, Been Kim, Fernanda Vi\u00e9gas, and Martin Wattenberg. 2017. Smoothgrad: removing noise by adding noise. arXiv preprint arXiv:1706.03825 (2017)."},{"key":"e_1_3_2_2_34_1","unstructured":"J Springenberg Alexey Dosovitskiy Thomas Brox and M Riedmiller. 2015. Striving for Simplicity: The All Convolutional Net. In ICLR (workshop track) .  J Springenberg Alexey Dosovitskiy Thomas Brox and M Riedmiller. 2015. Striving for Simplicity: The All Convolutional Net. In ICLR (workshop track) ."},{"key":"e_1_3_2_2_35_1","doi-asserted-by":"publisher","DOI":"10.5555\/3305890.3306024"},{"key":"e_1_3_2_2_36_1","volume-title":"Should Health Care Demand Interpretable Artificial Intelligence or Accept \"Black Box\" Medicine? Annals of Internal Medicine","author":"Wang Fei","year":"2019","unstructured":"Fei Wang , Rainu Kaushal , and Dhruv Khullar . 2019. Should Health Care Demand Interpretable Artificial Intelligence or Accept \"Black Box\" Medicine? Annals of Internal Medicine ( 2019 ). Fei Wang, Rainu Kaushal, and Dhruv Khullar. 2019. Should Health Care Demand Interpretable Artificial Intelligence or Accept \"Black Box\" Medicine? Annals of Internal Medicine (2019)."},{"key":"e_1_3_2_2_37_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/D19-1002"},{"key":"e_1_3_2_2_38_1","volume-title":"Fashion-mnist: a novel image dataset for benchmarking machine learning algorithms. arXiv preprint arXiv:1708.07747","author":"Xiao Han","year":"2017","unstructured":"Han Xiao , Kashif Rasul , and Roland Vollgraf . 2017. Fashion-mnist: a novel image dataset for benchmarking machine learning algorithms. arXiv preprint arXiv:1708.07747 ( 2017 ). Han Xiao, Kashif Rasul, and Roland Vollgraf. 2017. Fashion-mnist: a novel image dataset for benchmarking machine learning algorithms. arXiv preprint arXiv:1708.07747 (2017)."},{"key":"e_1_3_2_2_39_1","unstructured":"Chih-Kuan Yeh Cheng-Yu Hsieh Arun Suggala David I Inouye and Pradeep K Ravikumar. 2019. On the (In) fidelity and Sensitivity of Explanations. In Advances in Neural Information Processing Systems. 10965--10976.  Chih-Kuan Yeh Cheng-Yu Hsieh Arun Suggala David I Inouye and Pradeep K Ravikumar. 2019. On the (In) fidelity and Sensitivity of Explanations. In Advances in Neural Information Processing Systems. 10965--10976."},{"key":"e_1_3_2_2_40_1","volume-title":"Interpretable Deep Learning under Fire. arXiv preprint arXiv:1812.00891","author":"Zhang Xinyang","year":"2018","unstructured":"Xinyang Zhang , Ningfei Wang , Shouling Ji , Hua Shen , and Ting Wang . 2018. Interpretable Deep Learning under Fire. arXiv preprint arXiv:1812.00891 ( 2018 ). Xinyang Zhang, Ningfei Wang, Shouling Ji, Hua Shen, and Ting Wang. 2018. Interpretable Deep Learning under Fire. arXiv preprint arXiv:1812.00891 (2018)."},{"key":"e_1_3_2_2_41_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/P19-1045"}],"event":{"name":"KDD '20: The 26th ACM SIGKDD Conference on Knowledge Discovery and Data Mining","location":"Virtual Event CA USA","acronym":"KDD '20","sponsor":["SIGMOD ACM Special Interest Group on Management of Data","SIGKDD ACM Special Interest Group on Knowledge Discovery in Data"]},"container-title":["Proceedings of the 26th ACM SIGKDD International Conference on Knowledge Discovery &amp; Data Mining"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3394486.3403071","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3394486.3403071","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T22:41:38Z","timestamp":1750200098000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3394486.3403071"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,8,20]]},"references-count":41,"alternative-id":["10.1145\/3394486.3403071","10.1145\/3394486"],"URL":"https:\/\/doi.org\/10.1145\/3394486.3403071","relation":{},"subject":[],"published":{"date-parts":[[2020,8,20]]},"assertion":[{"value":"2020-08-20","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}