{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,4]],"date-time":"2026-05-04T13:01:55Z","timestamp":1777899715394,"version":"3.51.4"},"reference-count":220,"publisher":"Association for Computing Machinery (ACM)","issue":"11","funder":[{"name":"TCS Research Scholar program"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Comput. Surv."],"published-print":{"date-parts":[[2026,8,30]]},"abstract":"<jats:p>Deep Learning and Machine Learning based models have become extremely popular in text processing and information retrieval, owing to their remarkable effectiveness. However, the complex non-linear structures underlying these models make them largely inscrutable. A significant body of research has focused on increasing the transparency of these models. This article provides a broad overview of research on the explainability and interpretability of information retrieval methods, focusing primarily on both neural models for document ranking, and retrieval-augmented generation systems. A significant part of this research is inspired by, and strongly related to, similar work done in the area of natural language processing (NLP). For completeness, therefore, we also briefly review (in the Appendices) some recent studies from the NLP community on explaining word embeddings, sequence modeling, attention modules, transformers, and BERT. The concluding section suggests some possible directions for future research on this topic.<\/jats:p>","DOI":"10.1145\/3801957","type":"journal-article","created":{"date-parts":[[2026,3,26]],"date-time":"2026-03-26T20:56:43Z","timestamp":1774558603000},"page":"1-48","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["Explainability of Text Processing and Retrieval Methods: A Survey"],"prefix":"10.1145","volume":"58","author":[{"ORCID":"https:\/\/orcid.org\/0000-0001-8091-0685","authenticated-orcid":false,"given":"Sourav","family":"Saha","sequence":"first","affiliation":[{"name":"CVPR Unit, Indian Statistical Institute","place":["Kolkata, India"]}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0000-8823-9742","authenticated-orcid":false,"given":"Debapriyo","family":"Majumdar","sequence":"additional","affiliation":[{"name":"CVPR Unit, Indian Statistical Institute","place":["Kolkata, India"]}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-9045-9971","authenticated-orcid":false,"given":"Mandar","family":"Mitra","sequence":"additional","affiliation":[{"name":"CVPR Unit, Indian Statistical Institute","place":["Kolkata, India"]}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2026,5,1]]},"reference":[{"key":"e_1_3_3_2_2","first-page":"4190","volume-title":"Proc. ACL\u201920","author":"Abnar Samira","year":"2020","unstructured":"Samira Abnar and Willem Zuidema. 2020. Quantifying attention flow in transformers. In Proc. ACL\u201920. 4190\u20134197. Retrieved from https:\/\/aclanthology.org\/2020.acl-main.385"},{"key":"e_1_3_3_3_2","doi-asserted-by":"crossref","first-page":"681","DOI":"10.1162\/tacl_a_00667","article-title":"Evaluating correctness and faithfulness of instruction-following models for question answering","volume":"12","author":"Adlakha Vaibhav","year":"2024","unstructured":"Vaibhav Adlakha, Parishad BehnamGhader, Xing Han Lu, Nicholas Meade, and Siva Reddy. 2024. Evaluating correctness and faithfulness of instruction-following models for question answering. Transactions of the Association for Computational Linguistics 12 (2024), 681\u2013699. Retrieved from https:\/\/aclanthology.org\/2024.tacl-1.38\/","journal-title":"Transactions of the Association for Computational Linguistics"},{"key":"e_1_3_3_4_2","first-page":"3490","volume-title":"Proc. EMNLP-IJCNLP\u201919","author":"Yilmaz Zeynep Akkalyoncu","year":"2019","unstructured":"Zeynep Akkalyoncu Yilmaz, Wei Yang, Haotian Zhang, and Jimmy Lin. 2019. Cross-domain modeling of sentence-level evidence for document retrieval. In Proc. EMNLP-IJCNLP\u201919. 3490\u20133496. Retrieved from https:\/\/aclanthology.org\/D19-1352"},{"key":"e_1_3_3_5_2","unstructured":"Guillaume Alain and Yoshua Bengio. 2017. Understanding intermediate layers using linear classifier probes. (2017). arXiv:1610.01644. Retrieved from https:\/\/arxiv.org\/abs\/1610.01644"},{"key":"e_1_3_3_6_2","unstructured":"Avishek Anand Lijun Lyu Maximilian Idahl Yumeng Wang Jonas Wallat and Zijian Zhang. 2022. Explainable information retrieval: A survey. (2022). arXiv:2211.02405. Retrieved from https:\/\/arxiv.org\/abs\/2211.02405"},{"issue":"6671","key":"e_1_3_3_7_2","doi-asserted-by":"crossref","first-page":"669","DOI":"10.1126\/science.adi6000","article-title":"Prediction-powered inference","volume":"382","author":"Angelopoulos A. N.","year":"2023","unstructured":"A. N. Angelopoulos, S. Bates, C. Fannjiang, M. I. Jordan, and T. Zrnic. 2023. Prediction-powered inference. Science 382, 6671 (2023), 669\u2013674. Retrieved from https:\/\/www.science.org\/doi\/abs\/10.1126\/science.adi6000","journal-title":"Science"},{"key":"e_1_3_3_8_2","unstructured":"Kushal Arora Timothy J. O\u2019Donnell Doina Precup Jason Weston and Jackie C. K. Cheung. 2023. The stable entropy hypothesis and entropy-aware decoding: An analysis and algorithm for robust natural language generation. (2023) arxiv:2302.06784. Retrieved from https:\/\/arxiv.org\/abs\/2302.06784"},{"key":"e_1_3_3_9_2","volume-title":"Proc. ICLR\u201915","author":"Bahdanau Dzmitry","year":"2015","unstructured":"Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio. 2015. Neural machine translation by jointly learning to align and translate. In Proc. ICLR\u201915. arXiv:1409.0473. Retrieved from http:\/\/arxiv.org\/abs\/1409.0473"},{"key":"e_1_3_3_10_2","first-page":"14724","volume-title":"Findings ACL EMNLP 2024","author":"Balde Gunjan","year":"2024","unstructured":"Gunjan Balde, Soumyadeep Roy, Mainack Mondal, and Niloy Ganguly. 2024. Adaptive BPE tokenization for enhanced vocabulary adaptation in finetuning pretrained language models. In Findings ACL EMNLP 2024. 14724\u201314733. Retrieved from https:\/\/aclanthology.org\/2024.findings-emnlp.863"},{"key":"e_1_3_3_11_2","doi-asserted-by":"publisher","DOI":"10.1162\/coli_a_00422"},{"key":"e_1_3_3_12_2","doi-asserted-by":"crossref","first-page":"49","DOI":"10.1162\/tacl_a_00254","article-title":"Analysis methods in neural language processing: A survey","volume":"7","author":"Belinkov Yonatan","year":"2019","unstructured":"Yonatan Belinkov and James Glass. 2019. Analysis methods in neural language processing: A survey. TACL 7 (2019), 49\u201372. Retrieved from https:\/\/aclanthology.org\/Q19-1004","journal-title":"TACL"},{"key":"e_1_3_3_13_2","first-page":"14","volume-title":"Proc. ACL\u201918","author":"Blevins Terra","year":"2018","unstructured":"Terra Blevins, Omer Levy, and Luke Zettlemoyer. 2018. Deep RNNs encode soft hierarchical syntax. In Proc. ACL\u201918. 14\u201319. Retrieved from https:\/\/aclanthology.org\/P18-2003"},{"key":"e_1_3_3_14_2","unstructured":"Bernd Bohnet Vinh Q. Tran Pat Verga Roee Aharoni Daniel Andor Livio Baldini Soares Massimiliano Ciaramita Jacob Eisenstein Kuzman Ganchev Jonathan Herzig et\u00a0al. 2023. Attributed question answering: Evaluation and modeling for attributed large language models. (2023). arxiv:2212.08037. Retrieved from https:\/\/arxiv.org\/abs\/2212.08037"},{"key":"e_1_3_3_15_2","first-page":"632","volume-title":"Proc. EMNLP\u201915","author":"Bowman Samuel R.","year":"2015","unstructured":"Samuel R. Bowman, Gabor Angeli, Christopher Potts, and Christopher D. Manning. 2015. A large annotated corpus for learning natural language inference. In Proc. EMNLP\u201915. 632\u2013642. Retrieved from https:\/\/aclanthology.org\/D15-1075"},{"key":"e_1_3_3_16_2","unstructured":"Aakriti Budhraja Madhura Pande Pratyush Kumar and Mitesh M. Khapra. 2021. On the prunability of attention heads in multilingual BERT. (2021). arXiv:2109.12683. Retrieved from https:\/\/arxiv.org\/abs\/2109.12683"},{"key":"e_1_3_3_17_2","first-page":"25","volume-title":"Proc. Learning to Rank Challenge","author":"Burges Christopher","year":"2011","unstructured":"Christopher Burges, Krysta Svore, Paul Bennett, Andrzej Pastusiak, and Qiang Wu. 2011. Learning to rank using an ensemble of lambda-gradient models. In Proc. Learning to Rank Challenge, Vol. 14. 25\u201335. Retrieved from https:\/\/proceedings.mlr.press\/v14\/burges11a.html"},{"key":"e_1_3_3_18_2","unstructured":"Jamie Callan Mark Hoy Changkuk Yoo and Le Zhao. 2009. Clueweb09 data set. Retrieved from https:\/\/lemurproject.org\/clueweb09.php\/"},{"key":"e_1_3_3_19_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-45439-5_40"},{"key":"e_1_3_3_20_2","doi-asserted-by":"publisher","DOI":"10.1145\/290941.291025"},{"key":"e_1_3_3_21_2","series-title":"Proc. of Machine Learning Research","first-page":"1","volume-title":"Proc. Learning to Rank Challenge","author":"Chapelle Olivier","year":"2011","unstructured":"Olivier Chapelle and Yi Chang. 2011. Yahoo! learning to rank challenge overview. In Proc. Learning to Rank Challenge(Proc. of Machine Learning Research, Vol. 14). 1\u201324. Retrieved from https:\/\/proceedings.mlr.press\/v14\/chapelle11a.html"},{"key":"e_1_3_3_22_2","first-page":"5578","volume-title":"Proc. ACL\u201920","author":"Chen Hanjie","year":"2020","unstructured":"Hanjie Chen, Guangtao Zheng, and Yangfeng Ji. 2020. Generating hierarchical explanations on text classification via feature interaction detection. In Proc. ACL\u201920. Online, 5578\u20135593. Retrieved from https:\/\/aclanthology.org\/2020.acl-main.494"},{"key":"e_1_3_3_23_2","series-title":"ICTIR \u201925","doi-asserted-by":"crossref","first-page":"336","DOI":"10.1145\/3731120.3744603","volume-title":"Proceedings of the 2025 International ACM SIGIR Conference on Innovative Concepts and Theories in Information Retrieval (ICTIR)","author":"Chowdhury Tanya","year":"2025","unstructured":"Tanya Chowdhury, Atharva Nijasure, and James Allan. 2025. Probing ranking LLMs: A mechanistic analysis for information retrieval. In Proceedings of the 2025 International ACM SIGIR Conference on Innovative Concepts and Theories in Information Retrieval (ICTIR) (Padua, Italy) (ICTIR \u201925). Association for Computing Machinery, New York, NY, USA, 336\u2013346. DOI:10.1145\/3731120.3744603"},{"key":"e_1_3_3_24_2","volume-title":"Proceedings of the 13th International Conference on Learning Representations, ICLR 2025, Singapore, April 24-28, 2025","author":"Chowdhury Tanya","year":"2025","unstructured":"Tanya Chowdhury, Yair Zick, and James Allan. 2025. RankSHAP: Shapley value based feature attributions for learning to rank. In Proceedings of the 13th International Conference on Learning Representations, ICLR 2025, Singapore, April 24-28, 2025. OpenReview.net. Retrieved from https:\/\/openreview.net\/forum?id=4011PUI9vm"},{"key":"e_1_3_3_25_2","first-page":"8189","volume-title":"Proc. EMNLP\u201921","author":"Chrysostomou George","year":"2021","unstructured":"George Chrysostomou and Nikolaos Aletras. 2021. Enjoy the salience: Towards better transformer-based faithful explanations with word salience. In Proc. EMNLP\u201921. 8189\u20138200. Retrieved from https:\/\/aclanthology.org\/2021.emnlp-main.645"},{"key":"e_1_3_3_26_2","first-page":"276","volume-title":"Proc. ACL\u201919 Workshop BlackboxNLP: Analyzing and Interpreting Neural Networks for NLP","author":"Clark Kevin","year":"2019","unstructured":"Kevin Clark, Urvashi Khandelwal, Omer Levy, and Christopher D. Manning. 2019. What does BERT look at? an analysis of BERT\u2019s attention. In Proc. ACL\u201919 Workshop BlackboxNLP: Analyzing and Interpreting Neural Networks for NLP. 276\u2013286. Retrieved from https:\/\/aclanthology.org\/W19-4828"},{"key":"e_1_3_3_27_2","doi-asserted-by":"crossref","first-page":"78","DOI":"10.3390\/make5010006","article-title":"XAIR: A systematic metareview of explainable AI (XAI) aligned to the software development process","volume":"5","author":"Clement Tobias","year":"2023","unstructured":"Tobias Clement, Nils Kemmerzell, Mohamed Abdelaal, and Michael Amberg. 2023. XAIR: A systematic metareview of explainable AI (XAI) aligned to the software development process. Machine Learning and Knowledge Extraction 5, 1 (2023), 78\u2013108. Retrieved from https:\/\/api.semanticscholar.org\/CorpusID:255902137","journal-title":"Machine Learning and Knowledge Extraction"},{"key":"e_1_3_3_28_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-04417-5_6"},{"key":"e_1_3_3_29_2","doi-asserted-by":"publisher","DOI":"10.1145\/1645953.1646280"},{"key":"e_1_3_3_30_2","doi-asserted-by":"publisher","DOI":"10.1145\/2499178.2499179"},{"key":"e_1_3_3_31_2","doi-asserted-by":"crossref","first-page":"370","DOI":"10.18653\/v1\/2024.acl-short.35","volume-title":"Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers)","author":"Coelho Jo\u00e3o","year":"2024","unstructured":"Jo\u00e3o Coelho, Bruno Martins, Joao Magalhaes, Jamie Callan, and Chenyan Xiong. 2024. Dwell in the beginning: How language models embed long documents for dense retrieval. In Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers), Lun-Wei Ku, Andre Martins, and Vivek Srikumar (Eds.). Association for Computational Linguistics, Bangkok, Thailand, 370\u2013377. Retrieved from https:\/\/aclanthology.org\/2024.acl-short.35\/"},{"key":"e_1_3_3_32_2","doi-asserted-by":"publisher","DOI":"10.1145\/3234944.3234959"},{"key":"e_1_3_3_33_2","doi-asserted-by":"publisher","DOI":"10.1145\/3209978.3210118"},{"key":"e_1_3_3_34_2","volume-title":"Proc. LREC\u201918","author":"Conneau Alexis","year":"2018","unstructured":"Alexis Conneau and Douwe Kiela. 2018. SentEval: An evaluation toolkit for universal sentence representations. In Proc. LREC\u201918. Retrieved from https:\/\/aclanthology.org\/L18-1269"},{"key":"e_1_3_3_35_2","first-page":"2126","volume-title":"Proc. ACL\u201918","author":"Conneau Alexis","year":"2018","unstructured":"Alexis Conneau, German Kruszewski, Guillaume Lample, Lo\u00efc Barrault, and Marco Baroni. 2018. What you can cram into a single $&!#* vector: Probing sentence embeddings for linguistic properties. In Proc. ACL\u201918. 2126\u20132136. Retrieved from https:\/\/aclanthology.org\/P18-1198"},{"key":"e_1_3_3_36_2","volume-title":"Proc. TREC 2020","author":"Craswell Nick","year":"2021","unstructured":"Nick Craswell, Bhaskar Mitra, Emine Yilmaz, and Daniel Campos. 2021. Overview of the TREC 2020 deep learning track. In Proc. TREC 2020. NIST Special Publication: NIST SP 1266. arXiv:2102.07662. Retrieved from https:\/\/arxiv.org\/abs\/2102.07662"},{"key":"e_1_3_3_37_2","volume-title":"Proc. TREC 2019","author":"Craswell Nick","year":"2020","unstructured":"Nick Craswell, Bhaskar Mitra, Emine Yilmaz, Daniel Campos, and Ellen M. Voorhees. 2020. Overview of the TREC 2019 deep learning track. In Proc. TREC 2019. NIST Special Publication: SP 500-331. arXiv:2003.07820. Retrieved from https:\/\/arxiv.org\/abs\/2003.07820"},{"key":"e_1_3_3_38_2","doi-asserted-by":"publisher","DOI":"10.1145\/3331184.3331303"},{"key":"e_1_3_3_39_2","first-page":"447","volume-title":"Proc. AACL\u201920","author":"Danilevsky Marina","year":"2020","unstructured":"Marina Danilevsky, Kun Qian, Ranit Aharonov, Yannis Katsis, Ban Kawas, and Prithviraj Sen. 2020. A survey of the state of explainable AI for natural language processing. In Proc. AACL\u201920. 447\u2013459. Retrieved from https:\/\/aclanthology.org\/2020.aacl-main.46"},{"key":"e_1_3_3_40_2","first-page":"4171","volume-title":"Proc. NAACL-HLT\u201919","author":"Devlin Jacob","year":"2019","unstructured":"Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019. BERT: Pre-training of deep bidirectional transformers for language understanding. In Proc. NAACL-HLT\u201919. 4171\u20134186. Retrieved from https:\/\/aclanthology.org\/N19-1423"},{"key":"e_1_3_3_41_2","first-page":"4443","volume-title":"Proc. ACL\u201920","author":"DeYoung Jay","year":"2020","unstructured":"Jay DeYoung, Sarthak Jain, Nazneen Fatema Rajani, Eric Lehman, Caiming Xiong, Richard Socher, and Byron C. Wallace. 2020. ERASER: A benchmark to evaluate rationalized NLP models. In Proc. ACL\u201920. 4443\u20134458. Retrieved from https:\/\/aclanthology.org\/2020.acl-main.408"},{"key":"e_1_3_3_42_2","first-page":"367","volume-title":"Proc. ACL\u201916","author":"Diaz Fernando","year":"2016","unstructured":"Fernando Diaz, Bhaskar Mitra, and Nick Craswell. 2016. Query expansion with locally-trained word embeddings. In Proc. ACL\u201916. 367\u2013377. Retrieved from https:\/\/aclanthology.org\/P16-1035"},{"key":"e_1_3_3_43_2","unstructured":"Emily Dinan Stephen Roller Kurt Shuster Angela Fan Michael Auli and Jason Weston. 2018. Wizard of wikipedia: Knowledge-powered conversational agents. (2018). arXiv:1811.01241. Retrieved from http:\/\/arxiv.org\/abs\/1811.01241"},{"key":"e_1_3_3_44_2","volume-title":"Proc. IWP\u201905","author":"Dolan William B.","year":"2005","unstructured":"William B. Dolan and Chris Brockett. 2005. Automatically constructing a corpus of sentential paraphrases. In Proc. IWP\u201905. Retrieved from https:\/\/aclanthology.org\/I05-5002"},{"key":"e_1_3_3_45_2","unstructured":"Finale Doshi-Velez and Been Kim. 2017. Towards a rigorous science of interpretable machine learning. (2017). arXiv:1702.08608. Retrieved from http:\/\/arxiv.org\/abs\/1702.08608"},{"key":"e_1_3_3_46_2","doi-asserted-by":"publisher","DOI":"10.1162\/coli_a_00445"},{"key":"e_1_3_3_47_2","first-page":"1185","volume-title":"Proc. EMNLP-IJCNLP\u201919","author":"Dufter Philipp","year":"2019","unstructured":"Philipp Dufter and Hinrich Sch\u00fctze. 2019. Analytical methods for interpretable ultradense word embeddings. In Proc. EMNLP-IJCNLP\u201919. 1185\u20131191. Retrieved from https:\/\/aclanthology.org\/D19-1111"},{"key":"e_1_3_3_48_2","first-page":"150","volume-title":"Proceedings of the 18th Conference of the European Chapter of the Association for Computational Linguistics: System Demonstrations","author":"Es Shahul","year":"2024","unstructured":"Shahul Es, Jithin James, Luis Espinosa Anke, and Steven Schockaert. 2024. RAGAs: Automated evaluation of retrieval augmented generation. In Proceedings of the 18th Conference of the European Chapter of the Association for Computational Linguistics: System Demonstrations, Nikolaos Aletras and Orphee De Clercq (Eds.). Association for Computational Linguistics, St. Julians, Malta, 150\u2013158. Retrieved from https:\/\/aclanthology.org\/2024.eacl-demo.16\/"},{"key":"e_1_3_3_49_2","doi-asserted-by":"crossref","first-page":"34","DOI":"10.1162\/tacl_a_00298","article-title":"What BERT is not: Lessons from a new suite of psycholinguistic diagnostics for language models","volume":"8","author":"Ettinger Allyson","year":"2020","unstructured":"Allyson Ettinger. 2020. What BERT is not: Lessons from a new suite of psycholinguistic diagnostics for language models. TACL 8 (2020), 34\u201348. Retrieved from https:\/\/aclanthology.org\/2020.tacl-1.3","journal-title":"TACL"},{"key":"e_1_3_3_50_2","doi-asserted-by":"publisher","DOI":"10.1145\/1008992.1009004"},{"key":"e_1_3_3_51_2","doi-asserted-by":"publisher","DOI":"10.1145\/1961209.1961210"},{"key":"e_1_3_3_52_2","doi-asserted-by":"publisher","DOI":"10.1145\/1148170.1148193"},{"key":"e_1_3_3_53_2","doi-asserted-by":"publisher","DOI":"10.1145\/3331184.3331312"},{"key":"e_1_3_3_54_2","first-page":"257","volume-title":"Proc. ECIR\u201921","author":"Formal Thibault","year":"2021","unstructured":"Thibault Formal, Benjamin Piwowarski, and St\u00e9phane Clinchant. 2021. A white box analysis of ColBERT. In Proc. ECIR\u201921. 257\u2013263. Retrieved from https:\/\/hal.sorbonne-universite.fr\/hal-03364396"},{"key":"e_1_3_3_55_2","doi-asserted-by":"publisher","DOI":"10.1162\/tacl_a_00370"},{"key":"e_1_3_3_56_2","unstructured":"Yoav Goldberg. 2019. Assessing BERT\u2019s syntactic abilities. (2019). arXiv:1901.05287. Retrieved from http:\/\/arxiv.org\/abs\/1901.05287"},{"key":"e_1_3_3_57_2","doi-asserted-by":"publisher","unstructured":"Maarten Grootendorst. 2020. KeyBERT: Minimal keyword extraction with BERT. DOI:10.5281\/zenodo.4461265","DOI":"10.5281\/zenodo.4461265"},{"key":"e_1_3_3_58_2","doi-asserted-by":"publisher","DOI":"10.1145\/2983323.2983769"},{"key":"e_1_3_3_59_2","first-page":"12","volume-title":"Proc. EMNLP\u201915","author":"Gupta Abhijeet","year":"2015","unstructured":"Abhijeet Gupta, Gemma Boleda, Marco Baroni, and Sebastian Pad\u00f3. 2015. Distributional vectors encode referential attributes. In Proc. EMNLP\u201915. 12\u201321. Retrieved from https:\/\/aclanthology.org\/D15-1002"},{"key":"e_1_3_3_60_2","doi-asserted-by":"publisher","DOI":"10.1145\/2983323.2983704"},{"key":"e_1_3_3_61_2","volume-title":"Proc. AAAI\u201921","author":"Hao Yaru","year":"2021","unstructured":"Yaru Hao, Li Dong, Furu Wei, and Ke Xu. 2021. Self-attention attribution: Interpreting information interactions inside transformer. In Proc. AAAI\u201921. Retrieved from https:\/\/arxiv.org\/pdf\/2004.11207.pdf"},{"key":"e_1_3_3_62_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-45442-5_21"},{"key":"e_1_3_3_63_2","doi-asserted-by":"publisher","DOI":"10.1145\/3459637.3482445"},{"key":"e_1_3_3_64_2","doi-asserted-by":"publisher","DOI":"10.1214\/ss\/1177013604"},{"key":"e_1_3_3_65_2","doi-asserted-by":"publisher","DOI":"10.1145\/3726302.3730361"},{"key":"e_1_3_3_66_2","series-title":"SIGIR \u201925","doi-asserted-by":"crossref","first-page":"381","DOI":"10.1145\/3726302.3729971","volume-title":"Proceedings of the 48th International ACM SIGIR Conference on Research and Development in Information Retrieval","author":"Heuss Maria","year":"2025","unstructured":"Maria Heuss, Maarten de Rijke, and Avishek Anand. 2025. RankingSHAP - Faithful listwise feature attribution explanations for ranking models. In Proceedings of the 48th International ACM SIGIR Conference on Research and Development in Information Retrieval (Padua, Italy) (SIGIR \u201925). Association for Computing Machinery, New York, NY, USA, 381\u2013391. DOI:10.1145\/3726302.3729971"},{"key":"e_1_3_3_67_2","doi-asserted-by":"publisher","DOI":"10.1162\/neco.1997.9.8.1735"},{"key":"e_1_3_3_68_2","first-page":"2042","volume-title":"Proc. NIPS\u201914","author":"Hu Baotian","year":"2014","unstructured":"Baotian Hu, Zhengdong Lu, Hang Li, and Qingcai Chen. 2014. Convolutional neural network architectures for matching natural language sentences. In Proc. NIPS\u201914. 2042\u20132050. Retrieved from https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2014\/file\/b9d487a30398d42ecff55c228ed5652b-Paper.pdf"},{"key":"e_1_3_3_69_2","volume-title":"Proceedings of the International Conference on Learning Representations","author":"Hu Edward J.","year":"2022","unstructured":"Edward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. 2022. LoRA: Low-rank adaptation of large language models. In Proceedings of the International Conference on Learning Representations. Retrieved from https:\/\/openreview.net\/forum?id=nZeVKeeFYf9"},{"key":"e_1_3_3_70_2","doi-asserted-by":"publisher","DOI":"10.1613\/jair.1.11196"},{"key":"e_1_3_3_71_2","article-title":"Unsupervised dense information retrieval with contrastive learning.","volume":"2022","author":"Izacard Gautier","year":"2022","unstructured":"Gautier Izacard, Mathilde Caron, Lucas Hosseini, Sebastian Riedel, Piotr Bojanowski, Armand Joulin, and Edouard Grave. 2022. Unsupervised dense information retrieval with contrastive learning. Transactions on Machine Learning Research 2022 (2022). Retrieved from http:\/\/dblp.uni-trier.de\/db\/journals\/tmlr\/tmlr2022.html#IzacardCHRBJG22","journal-title":"Transactions on Machine Learning Research"},{"key":"e_1_3_3_72_2","doi-asserted-by":"crossref","first-page":"294","DOI":"10.1162\/tacl_a_00367","article-title":"Aligning faithful interpretations with their social attribution","volume":"9","author":"Jacovi Alon","year":"2021","unstructured":"Alon Jacovi and Yoav Goldberg. 2021. Aligning faithful interpretations with their social attribution. Transactions of the Association for Computational Linguistics 9 (2021), 294\u2013310. Retrieved from https:\/\/aclanthology.org\/2021.tacl-1.18\/","journal-title":"Transactions of the Association for Computational Linguistics"},{"key":"e_1_3_3_73_2","doi-asserted-by":"crossref","first-page":"1597","DOI":"10.18653\/v1\/2021.emnlp-main.120","volume-title":"Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing","author":"Jacovi Alon","year":"2021","unstructured":"Alon Jacovi, Swabha Swayamdipta, Shauli Ravfogel, Yanai Elazar, Yejin Choi, and Yoav Goldberg. 2021. Contrastive explanations for model interpretability. In Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing, Marie-Francine Moens, Xuanjing Huang, Lucia Specia, and Scott Wen-tau Yih (Eds.). Association for Computational Linguistics, Online and Punta Cana, Dominican Republic, 1597\u20131611. DOI:10.18653\/v1\/2021.emnlp-main.120"},{"key":"e_1_3_3_74_2","first-page":"3543","volume-title":"Proc. NAACL-HLT\u201919","author":"Jain Sarthak","year":"2019","unstructured":"Sarthak Jain and Byron C. Wallace. 2019. Attention is not explanation. In Proc. NAACL-HLT\u201919. 3543\u20133556. Retrieved from https:\/\/aclanthology.org\/N19-1357"},{"key":"e_1_3_3_75_2","first-page":"3651","volume-title":"Proc. ACL\u201919","author":"Jawahar Ganesh","year":"2019","unstructured":"Ganesh Jawahar, Beno\u00eet Sagot, and Djam\u00e9 Seddah. 2019. What does BERT learn about the structure of language?. In Proc. ACL\u201919. 3651\u20133657. Retrieved from https:\/\/aclanthology.org\/P19-1356"},{"key":"e_1_3_3_76_2","doi-asserted-by":"publisher","DOI":"10.1007\/s10791-005-0750-7"},{"key":"e_1_3_3_77_2","first-page":"6769","volume-title":"Proc. EMNLP\u201920","author":"Karpukhin Vladimir","year":"2020","unstructured":"Vladimir Karpukhin, Barlas Oguz, Sewon Min, Patrick Lewis, Ledell Wu, Sergey Edunov, Danqi Chen, and Wen-tau Yih. 2020. Dense passage retrieval for open-domain question answering. In Proc. EMNLP\u201920. Online, 6769\u20136781. Retrieved from https:\/\/aclanthology.org\/2020.emnlp-main.550"},{"key":"e_1_3_3_78_2","first-page":"3149","volume-title":"Proc. NIPS\u201917","author":"Ke Guolin","year":"2017","unstructured":"Guolin Ke, Qi Meng, Thomas Finley, Taifeng Wang, Wei Chen, Weidong Ma, Qiwei Ye, and Tie-Yan Liu. 2017. LightGBM: A highly efficient gradient boosting decision tree. In Proc. NIPS\u201917. 3149\u20133157. Retrieved from https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2017\/file\/6449f44a102fde848669bdd9eb6b76fa-Paper.pdf"},{"key":"e_1_3_3_79_2","first-page":"252","volume-title":"Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long Papers)","author":"Khashabi Daniel","year":"2018","unstructured":"Daniel Khashabi, Snigdha Chaturvedi, Michael Roth, Shyam Upadhyay, and Dan Roth. 2018. Looking beyond the surface: A challenge set for reading comprehension over multiple sentences. In Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long Papers). 252\u2013262. DOI:10.18653\/v1\/N18-1023"},{"key":"e_1_3_3_80_2","doi-asserted-by":"publisher","DOI":"10.1145\/3397271.3401075"},{"key":"e_1_3_3_81_2","doi-asserted-by":"publisher","DOI":"10.1145\/3477495.3531883"},{"key":"e_1_3_3_82_2","first-page":"2067","volume-title":"Proc. EMNLP\u201915","author":"K\u00f6hn Arne","year":"2015","unstructured":"Arne K\u00f6hn. 2015. What\u2019s in an embedding? analyzing word embeddings through multilingual evaluation. In Proc. EMNLP\u201915. 2067\u20132073. Retrieved from https:\/\/aclanthology.org\/D15-1246"},{"key":"e_1_3_3_83_2","first-page":"4365","volume-title":"Proc. EMNLP-IJCNLP\u201919","author":"Kovaleva Olga","year":"2019","unstructured":"Olga Kovaleva, Alexey Romanov, Anna Rogers, and Anna Rumshisky. 2019. Revealing the dark secrets of BERT. In Proc. EMNLP-IJCNLP\u201919. 4365\u20134374. Retrieved from https:\/\/aclanthology.org\/D19-1445"},{"key":"e_1_3_3_84_2","first-page":"957","volume-title":"Proc. ICML\u201915","author":"Kusner Matt J.","year":"2015","unstructured":"Matt J. Kusner, Yu Sun, Nicholas I. Kolkin, and Kilian Q. Weinberger. 2015. From word embeddings to document distances. In Proc. ICML\u201915. 957\u2013966. Retrieved from https:\/\/proceedings.mlr.press\/v37\/kusnerb15.html"},{"key":"e_1_3_3_85_2","first-page":"452","article-title":"Natural questions: A benchmark for question answering research","volume":"7","author":"Kwiatkowski Tom","year":"2019","unstructured":"Tom Kwiatkowski, Jennimaria Palomaki, Olivia Redfield, Michael Collins, Ankur Parikh, Chris Alberti, Danielle Epstein, Illia Polosukhin, Jacob Devlin, Kenton Lee, et\u00a0al. 2019. Natural questions: A benchmark for question answering research. Transactions of the Association for Computational Linguistics 7 (2019), 452\u2013466. Retrieved from https:\/\/aclanthology.org\/Q19-1026\/","journal-title":"Transactions of the Association for Computational Linguistics"},{"key":"e_1_3_3_86_2","first-page":"6423","volume-title":"Proc. NIPS\u201918","author":"Laha Anirban","year":"2018","unstructured":"Anirban Laha, Saneem A. Chemmengath, Priyanka Agrawal, Mitesh M. Khapra, Karthik Sankaranarayanan, and Harish G. Ramaswamy. 2018. On controllable sparse alternatives to softmax. In Proc. NIPS\u201918. 6423\u20136433. Retrieved from https:\/\/dl.acm.org\/doi\/pdf\/10.5555\/3327345.3327538"},{"key":"e_1_3_3_87_2","series-title":"SIGIR \u201924","doi-asserted-by":"crossref","first-page":"1040","DOI":"10.1145\/3626772.3657768","volume-title":"Proceedings of the 47th International ACM SIGIR Conference on Research and Development in Information Retrieval","author":"\u0141ajewska Weronika","year":"2024","unstructured":"Weronika \u0141ajewska, Damiano Spina, Johanne Trippas, and Krisztian Balog. 2024. Explainability for transparent conversational information-seeking. In Proceedings of the 47th International ACM SIGIR Conference on Research and Development in Information Retrieval (Washington DC, USA) (SIGIR \u201924). Association for Computing Machinery, New York, NY, USA, 1040\u20131050. DOI:10.1145\/3626772.3657768"},{"key":"e_1_3_3_88_2","volume-title":"Proc. NeurIPS\u201919","author":"Laue Soeren","year":"2019","unstructured":"Soeren Laue, Matthias Mitterreiter, and Joachim Giesen. 2019. GENO \u2013 GENeric optimization for classical machine learning. In Proc. NeurIPS\u201919, Vol. 32. Retrieved from https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2019\/file\/84438b7aae55a0638073ef798e50b4ef-Paper.pdf"},{"key":"e_1_3_3_89_2","doi-asserted-by":"publisher","DOI":"10.1145\/383952.383972"},{"key":"e_1_3_3_90_2","doi-asserted-by":"publisher","DOI":"10.1145\/3688392"},{"key":"e_1_3_3_91_2","doi-asserted-by":"publisher","DOI":"10.1145\/3576924"},{"key":"e_1_3_3_92_2","doi-asserted-by":"publisher","DOI":"10.1162\/tacl_a_00440"},{"key":"e_1_3_3_93_2","first-page":"7871","volume-title":"Proc. ACL\u201920","author":"Lewis Mike","year":"2020","unstructured":"Mike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad, Abdelrahman Mohamed, Omer Levy, Veselin Stoyanov, and Luke Zettlemoyer. 2020. BART: Denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension. In Proc. ACL\u201920. 7871\u20137880. Retrieved from https:\/\/aclanthology.org\/2020.acl-main.703"},{"key":"e_1_3_3_94_2","series-title":"NIPS \u201920","volume-title":"Proceedings of the 34th International Conference on Neural Information Processing Systems","author":"Lewis Patrick","year":"2020","unstructured":"Patrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich K\u00fcttler, Mike Lewis, Wen-tau Yih, Tim Rockt\u00e4schel, et\u00a0al. 2020. Retrieval-augmented generation for knowledge-intensive NLP tasks. In Proceedings of the 34th International Conference on Neural Information Processing Systems (Vancouver, BC, Canada) (NIPS \u201920). Curran Associates Inc., Red Hook, NY, USA, Article 793, 16 pages."},{"key":"e_1_3_3_95_2","unstructured":"Jiwei Li Will Monroe and Dan Jurafsky. 2016. Understanding neural networks through representation erasure. (2016). arXiv:1612.08220. Retrieved from http:\/\/arxiv.org\/abs\/1612.08220"},{"key":"e_1_3_3_96_2","doi-asserted-by":"publisher","DOI":"10.2200\/S01123ED1V01Y202108HLT053"},{"key":"e_1_3_3_97_2","first-page":"241","volume-title":"Proc. ACL\u201919 Workshop BlackboxNLP","author":"Lin Yongjie","year":"2019","unstructured":"Yongjie Lin, Yi Chern Tan, and Robert Frank. 2019. Open sesame: Getting inside BERT\u2019s linguistic knowledge. In Proc. ACL\u201919 Workshop BlackboxNLP. 241\u2013253. Retrieved from https:\/\/aclanthology.org\/W19-4825"},{"key":"e_1_3_3_98_2","doi-asserted-by":"publisher","DOI":"10.1145\/3236386.3241340"},{"key":"e_1_3_3_99_2","doi-asserted-by":"crossref","first-page":"157","DOI":"10.1162\/tacl_a_00638","article-title":"Lost in the middle: How language models use long contexts","volume":"12","author":"Liu Nelson F.","year":"2024","unstructured":"Nelson F. Liu, Kevin Lin, John Hewitt, Ashwin Paranjape, Michele Bevilacqua, Fabio Petroni, and Percy Liang. 2024. Lost in the middle: How language models use long contexts. Transactions of the Association for Computational Linguistics 12 (2024), 157\u2013173. Retrieved from https:\/\/aclanthology.org\/2024.tacl-1.9\/","journal-title":"Transactions of the Association for Computational Linguistics"},{"key":"e_1_3_3_100_2","doi-asserted-by":"publisher","DOI":"10.1145\/3539618.3591777"},{"key":"e_1_3_3_101_2","doi-asserted-by":"publisher","DOI":"10.1145\/3539618.3591982"},{"key":"e_1_3_3_102_2","doi-asserted-by":"publisher","DOI":"10.1145\/2339530.2339556"},{"key":"e_1_3_3_103_2","doi-asserted-by":"publisher","DOI":"10.1145\/2487575.2487579"},{"key":"e_1_3_3_104_2","first-page":"4461","volume-title":"Proc. NeurIPS\u201921","author":"Lu Kaiji","year":"2021","unstructured":"Kaiji Lu, Zifan Wang, Piotr Mardziel, and Anupam Datta. 2021. Influence patterns for explaining information flow in BERT. In Proc. NeurIPS\u201921. 4461\u20134474. Retrieved from https:\/\/openreview.net\/forum?id=FYDE3I9fev0"},{"key":"e_1_3_3_105_2","doi-asserted-by":"publisher","DOI":"10.1145\/2911451.2914763"},{"key":"e_1_3_3_106_2","doi-asserted-by":"publisher","DOI":"10.1145\/3477495.3531840"},{"key":"e_1_3_3_107_2","unstructured":"Scott M. Lundberg Gabriel G. Erion and Su-In Lee. 2018. Consistent individualized feature attribution for tree ensembles. (2018). arXiv:1802.03888. Retrieved from https:\/\/arxiv.org\/abs\/1802.03888"},{"key":"e_1_3_3_108_2","first-page":"4768","volume-title":"Proc. NIPS\u201917","author":"Lundberg Scott M.","year":"2017","unstructured":"Scott M. Lundberg and Su-In Lee. 2017. A unified approach to interpreting model predictions. In Proc. NIPS\u201917. 4768\u20134777. Retrieved from https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2017\/file\/8a20a8621978632d76c43dfd28b67767-Paper.pdf"},{"key":"e_1_3_3_109_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-28244-7_41"},{"key":"e_1_3_3_110_2","first-page":"142","volume-title":"Proc. ACL\u201911: Human Language Technologies","author":"Maas Andrew L.","year":"2011","unstructured":"Andrew L. Maas, Raymond E. Daly, Peter T. Pham, Dan Huang, Andrew Y. Ng, and Christopher Potts. 2011. Learning word vectors for sentiment analysis. In Proc. ACL\u201911: Human Language Technologies. 142\u2013150. Retrieved from https:\/\/aclanthology.org\/P11-1015"},{"key":"e_1_3_3_111_2","doi-asserted-by":"publisher","DOI":"10.1162\/tacl_a_00457"},{"key":"e_1_3_3_112_2","doi-asserted-by":"publisher","DOI":"10.1145\/3397271.3401262"},{"key":"e_1_3_3_113_2","doi-asserted-by":"publisher","DOI":"10.1145\/3546577"},{"key":"e_1_3_3_114_2","first-page":"370","volume-title":"Proc. ACL\u201918","author":"Malaviya Chaitanya","year":"2018","unstructured":"Chaitanya Malaviya, Pedro Ferreira, and Andr\u00e9 F. T. Martins. 2018. Sparse and constrained attention for neural machine translation. In Proc. ACL\u201918. 370\u2013376. Retrieved from https:\/\/aclanthology.org\/P18-2059"},{"key":"e_1_3_3_115_2","doi-asserted-by":"crossref","first-page":"114","DOI":"10.3115\/1075812.1075835","volume-title":"Proceedings of the Workshop on Human Language Technology","author":"Marcus Mitchell","year":"1994","unstructured":"Mitchell Marcus, Grace Kim, Mary Ann Marcinkiewicz, Robert MacIntyre, Ann Bies, Mark Ferguson, Karen Katz, and Britta Schasberger. 1994. The Penn treebank: Annotating predicate argument structure. In Proceedings of the Workshop on Human Language Technology. 114\u2013119. DOI:10.3115\/1075812.1075835"},{"key":"e_1_3_3_116_2","first-page":"1614","volume-title":"Proc. ICML\u201916","author":"Martins Andre","year":"2016","unstructured":"Andre Martins and Ramon Astudillo. 2016. From softmax to sparsemax: A sparse model of attention and multi-label classification. In Proc. ICML\u201916, Vol. 48. 1614\u20131623. Retrieved from https:\/\/proceedings.mlr.press\/v48\/martins16.html"},{"key":"e_1_3_3_117_2","doi-asserted-by":"publisher","DOI":"10.1145\/3366423.3380227"},{"key":"e_1_3_3_118_2","first-page":"6297","volume-title":"Proc. NIPS\u201917","author":"McCann Bryan","year":"2017","unstructured":"Bryan McCann, James Bradbury, Caiming Xiong, and Richard Socher. 2017. Learned in translation: Contextualized word vectors. In Proc. NIPS\u201917. 6297\u20136308. Retrieved from https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2017\/file\/20c86a628232a67e7bd46f76fba7ce12-Paper.pdf"},{"key":"e_1_3_3_119_2","volume-title":"Proc. ICLR\u201919, New Orleans, LA, USA, May 6-9, 2019","author":"McCoy R. Thomas","year":"2019","unstructured":"R. Thomas McCoy, Tal Linzen, Ewan Dunbar, and Paul Smolensky. 2019. RNNs implicitly implement tensor-product representations. In Proc. ICLR\u201919, New Orleans, LA, USA, May 6-9, 2019. Retrieved from https:\/\/openreview.net\/forum?id=BJx0sjC5FX"},{"key":"e_1_3_3_120_2","first-page":"1849","volume-title":"Proc. EMNLP\u201918","author":"McDonald Ryan","year":"2018","unstructured":"Ryan McDonald, George Brokos, and Ion Androutsopoulos. 2018. Deep relevance ranking using enhanced document-query interactions. In Proc. EMNLP\u201918. 1849\u20131860. Retrieved from https:\/\/aclanthology.org\/D18-1211"},{"key":"e_1_3_3_121_2","first-page":"122","volume-title":"Proc. ACL-IJCNLP\u201921","author":"Meister Clara","year":"2021","unstructured":"Clara Meister, Stefan Lazov, Isabelle Augenstein, and Ryan Cotterell. 2021. Is sparse attention more interpretable?. In Proc. ACL-IJCNLP\u201921. 122\u2013129. Retrieved from https:\/\/aclanthology.org\/2021.acl-short.17"},{"key":"e_1_3_3_122_2","first-page":"404","volume-title":"Proc. EMNLP\u201904","author":"Mihalcea Rada","year":"2004","unstructured":"Rada Mihalcea and Paul Tarau. 2004. TextRank: Bringing order into text. In Proc. EMNLP\u201904. 404\u2013411. Retrieved from https:\/\/aclanthology.org\/W04-3252"},{"key":"e_1_3_3_123_2","first-page":"3111","volume-title":"Proc. NIPS\u201913","author":"Mikolov Tomas","year":"2013","unstructured":"Tomas Mikolov, Ilya Sutskever, Kai Chen, Greg Corrado, and Jeffrey Dean. 2013. Distributed representations of words and phrases and their compositionality. In Proc. NIPS\u201913 (Lake Tahoe, Nevada). 3111\u20133119. Retrieved from https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2013\/file\/9aa42b31882ec039965f3c4923ce901b-Paper.pdf"},{"key":"e_1_3_3_124_2","doi-asserted-by":"publisher","DOI":"10.1145\/3038912.3052579"},{"key":"e_1_3_3_125_2","first-page":"4206","volume-title":"Proc. ACL\u201920","author":"Mohankumar Akash Kumar","year":"2020","unstructured":"Akash Kumar Mohankumar, Preksha Nema, Sharan Narasimhan, Mitesh M. Khapra, Balaji Vasan Srinivasan, and Balaraman Ravindran. 2020. Towards transparent and explainable attention models. In Proc. ACL\u201920. 4206\u20134216. Retrieved from https:\/\/aclanthology.org\/2020.acl-main.387"},{"key":"e_1_3_3_126_2","first-page":"7055","volume-title":"Proc. ICML\u201920","author":"Moshkovitz Michal","year":"2020","unstructured":"Michal Moshkovitz, Sanjoy Dasgupta, Cyrus Rashtchian, and Nave Frost. 2020. Explainable k-Means and k-Medians Clustering. In Proc. ICML\u201920, Vol. 119. 7055\u20137065. Retrieved from https:\/\/proceedings.mlr.press\/v119\/moshkovitz20a.html"},{"issue":"19","key":"e_1_3_3_127_2","doi-asserted-by":"crossref","first-page":"3055","DOI":"10.1002\/sim.1545","article-title":"Estimating regression models with unknown break-points","volume":"22","author":"Muggeo Vito M. R.","year":"2003","unstructured":"Vito M. R. Muggeo. 2003. Estimating regression models with unknown break-points. Statistics in Medicine 22, 19 (2003), 3055\u20133071. Retrieved from https:\/\/onlinelibrary.wiley.com\/doi\/abs\/10.1002\/sim.1545","journal-title":"Statistics in Medicine"},{"key":"e_1_3_3_128_2","doi-asserted-by":"crossref","first-page":"144","DOI":"10.18653\/v1\/2023.emnlp-main.10","volume-title":"Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing","author":"Muller Benjamin","year":"2023","unstructured":"Benjamin Muller, John Wieting, Jonathan Clark, Tom Kwiatkowski, Sebastian Ruder, Livio Soares, Roee Aharoni, Jonathan Herzig, and Xinyi Wang. 2023. Evaluating and modeling attribution for cross-lingual question answering. In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing, Houda Bouamor, Juan Pino, and Kalika Bali (Eds.). Association for Computational Linguistics, Singapore, 144\u2013157. Retrieved from https:\/\/aclanthology.org\/2023.emnlp-main.10\/"},{"key":"e_1_3_3_129_2","volume-title":"Proc. ICLR\u201918","author":"Murdoch W. James","year":"2018","unstructured":"W. James Murdoch, Peter J. Liu, and Bin Yu. 2018. Beyond word importance: Contextual decomposition to extract interactions from LSTMs. In Proc. ICLR\u201918. Retrieved from https:\/\/openreview.net\/forum?id=rkRwGg-0Z"},{"key":"e_1_3_3_130_2","volume-title":"Proc. ICLR\u201917","author":"Murdoch W. James","year":"2017","unstructured":"W. James Murdoch and Arthur Szlam. 2017. Automatic rule extraction from long short term memory networks. In Proc. ICLR\u201917. OpenReview.net. Retrieved from https:\/\/openreview.net\/forum?id=SJvYgH9xe"},{"key":"e_1_3_3_131_2","doi-asserted-by":"publisher","DOI":"10.1145\/2872518.2889361"},{"key":"e_1_3_3_132_2","first-page":"1069","volume-title":"Proc. NAACL-HLT\u201918","author":"Nguyen Dong","year":"2018","unstructured":"Dong Nguyen. 2018. Comparing automatic and human evaluation of local explanations for text classification. In Proc. NAACL-HLT\u201918. 1069\u20131078. Retrieved from https:\/\/aclanthology.org\/N18-1097"},{"key":"e_1_3_3_133_2","unstructured":"Tri Nguyen Mir Rosenberg Xia Song Jianfeng Gao Saurabh Tiwary Rangan Majumder and Li Deng. 2016. MS MARCO: A human generated machine reading comprehension dataset. (2016). arXiv:1611.09268. Retrieved from http:\/\/arxiv.org\/abs\/1611.09268"},{"key":"e_1_3_3_134_2","unstructured":"Atharva Nijasure Tanya Chowdhury and James Allan. 2025. How Relevance Emerges: A Mechanistic Analysis of LoRA Fine-Tuning in Reranking LLMs. In Proceedings of the 2025 International ACM SIGIR Conference on Innovative Concepts and Theories in Information Retrieval (ICTIR) Pages 336-346."},{"key":"e_1_3_3_135_2","first-page":"4658","volume-title":"Proc. ACL\u201919","author":"Niven Timothy","year":"2019","unstructured":"Timothy Niven and Hung-Yu Kao. 2019. Probing neural network comprehension of natural language arguments. In Proc. ACL\u201919. Florence, Italy, 4658\u20134664. Retrieved from https:\/\/aclanthology.org\/P19-1459"},{"key":"e_1_3_3_136_2","unstructured":"Rodrigo Nogueira and Kyunghyun Cho. 2019. Passage Re-ranking with BERT. (2019). arXiv:1901.04085. Retrieved from http:\/\/arxiv.org\/abs\/1901.04085"},{"key":"e_1_3_3_137_2","unstructured":"Rodrigo Nogueira and Jimmy Lin. 2019. From doc2query to docTTTTTquery. Retrieved from https:\/\/cs.uwaterloo.ca\/jimmylin\/publications\/Nogueira_Lin_2019_docTTTTTquery-v2.pdf"},{"key":"e_1_3_3_138_2","volume-title":"Proceedings of the Text Retrieval Conference","author":"Owoicho Paul","year":"2022","unstructured":"Paul Owoicho, Jeffrey Dalton, Mohammad Aliannejadi, Leif Azzopardi, Johanne R. Trippas, and Svitlana Vakulenko. 2022. TREC CAsT 2022: Going beyond user ask and system retrieve with initiative and response generation. In Proceedings of the Text Retrieval Conference. Retrieved from https:\/\/api.semanticscholar.org\/CorpusID:261288646"},{"key":"e_1_3_3_139_2","doi-asserted-by":"publisher","DOI":"10.1109\/TASLP.2016.2520371"},{"key":"e_1_3_3_140_2","first-page":"13613","volume-title":"Proc. AAAI\/IAAI\/EAAI","author":"Pande Madhura","year":"2021","unstructured":"Madhura Pande, Aakriti Budhraja, Preksha Nema, Pratyush Kumar, and Mitesh M. Khapra. 2021. The heads hypothesis: A unifying statistical approach towards understanding multi-headed attention in BERT. In Proc. AAAI\/IAAI\/EAAI. 13613\u201313621. Retrieved from https:\/\/ojs.aaai.org\/index.php\/AAAI\/article\/view\/17605"},{"key":"e_1_3_3_141_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-56066-8_28"},{"key":"e_1_3_3_142_2","unstructured":"Liang Pang Yanyan Lan J. Guo Jun Xu and Xueqi Cheng. 2016. A study of MatchPyramid models on Ad-hoc retrieval. (2016). arXiv:1606.04648. Retrieved from https:\/\/arxiv.org\/abs\/1606.04648"},{"key":"e_1_3_3_143_2","unstructured":"Liang Pang Yanyan Lan J. Guo Jun Xu and Xueqi Cheng. 2017. A deep investigation of deep IR models. (2017). arXiv:1707.07700. Retrieved from https:\/\/arxiv.org\/abs\/1707.07700"},{"key":"e_1_3_3_144_2","first-page":"2793","volume-title":"Proc. AAAI\u201916","author":"Pang Liang","year":"2016","unstructured":"Liang Pang, Yanyan Lan, Jiafeng Guo, Jun Xu, Shengxian Wan, and Xueqi Cheng. 2016. Text matching as image recognition. In Proc. AAAI\u201916 (Phoenix, Arizona). 2793\u20132799. Retrieved from https:\/\/ojs.aaai.org\/index.php\/AAAI\/article\/view\/10341"},{"key":"e_1_3_3_145_2","series-title":"Proceedings of Machine Learning Research","first-page":"4574","volume-title":"Proceedings of the 25th International Conference on Artificial Intelligence and Statistics","volume":"151","author":"Pawelczyk Martin","year":"2022","unstructured":"Martin Pawelczyk, Chirag Agarwal, Shalmali Joshi, Sohini Upadhyay, and Himabindu Lakkaraju. 2022. Exploring counterfactual explanations through the lens of adversarial examples: A theoretical and empirical analysis. In Proceedings of the 25th International Conference on Artificial Intelligence and Statistics(Proceedings of Machine Learning Research, Vol. 151), Gustau Camps-Valls, Francisco J. R. Ruiz, and Isabel Valera (Eds.). PMLR, 4574\u20134594. Retrieved from https:\/\/proceedings.mlr.press\/v151\/pawelczyk22a.html"},{"key":"e_1_3_3_146_2","first-page":"1532","volume-title":"Proc. EMNLP\u201914","author":"Pennington Jeffrey","year":"2014","unstructured":"Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014. GloVe: Global vectors for word representation. In Proc. EMNLP\u201914. 1532\u20131543. Retrieved from https:\/\/aclanthology.org\/D14-1162"},{"key":"e_1_3_3_147_2","first-page":"2227","volume-title":"Proc. NAACL-HLT\u201918","author":"Peters Matthew E.","year":"2018","unstructured":"Matthew E. Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer. 2018. Deep contextualized word representations. In Proc. NAACL-HLT\u201918. 2227\u20132237. Retrieved from https:\/\/aclanthology.org\/N18-1202"},{"key":"e_1_3_3_148_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-99739-7_65"},{"key":"e_1_3_3_149_2","first-page":"6037","volume-title":"Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing","author":"Qi Jirui","year":"2024","unstructured":"Jirui Qi, Gabriele Sarti, Raquel Fern\u00e1ndez, and Arianna Bisazza. 2024. Model internals-based answer attribution for trustworthy retrieval-augmented generation. In Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing, Yaser Al-Onaizan, Mohit Bansal, and Yun-Nung Chen (Eds.). Association for Computational Linguistics, Miami, Florida, USA, 6037\u20136053. Retrieved from https:\/\/aclanthology.org\/2024.emnlp-main.347\/"},{"key":"e_1_3_3_150_2","unstructured":"Yifan Qiao Chenyan Xiong Zhenghao Liu and Zhiyuan Liu. 2019. Understanding the behaviors of BERT in ranking. (2019). arXiv:1904.07531. Retrieved from https:\/\/arxiv.org\/abs\/1904.07531"},{"key":"e_1_3_3_151_2","unstructured":"Tao Qin and Tie-Yan Liu. 2013. Introducing LETOR 4.0 datasets. (2013). arXiv:1306.2597. Retrieved from https:\/\/arxiv.org\/abs\/1306.2597"},{"key":"e_1_3_3_152_2","doi-asserted-by":"publisher","DOI":"10.1007\/s10791-009-9123-y"},{"key":"e_1_3_3_153_2","unstructured":"Alec Radford and Karthik Narasimhan. 2018. Improving Language Understanding by Generative Pre- Training. Retrieved from https:\/\/cdn.openai.com\/research-covers\/language-unsupervised\/language_understanding_paper.pdf"},{"issue":"140","key":"e_1_3_3_154_2","first-page":"1","article-title":"Exploring the limits of transfer learning with a unified text-to-text transformer","volume":"21","author":"Raffel Colin","year":"2020","unstructured":"Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu. 2020. Exploring the limits of transfer learning with a unified text-to-text transformer. Journal of Machine Learning Research 21, 140 (2020), 1\u201367. Retrieved from http:\/\/jmlr.org\/papers\/v21\/20-074.html","journal-title":"Journal of Machine Learning Research"},{"key":"e_1_3_3_155_2","unstructured":"Razieh Rahimi Youngwoo Kim Hamed Zamani and James Allan. 2021. Explaining documents\u2019 relevance to search queries. (2021). arXiv:2111.01314. Retrieved from https:\/\/arxiv.org\/abs\/2111.01314"},{"key":"e_1_3_3_156_2","doi-asserted-by":"crossref","first-page":"2383","DOI":"10.18653\/v1\/D16-1264","volume-title":"Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing","author":"Rajpurkar Pranav","year":"2016","unstructured":"Pranav Rajpurkar, Jian Zhang, Konstantin Lopyrev, and Percy Liang. 2016. SQuAD: 100,000+ questions for machine comprehension of text. In Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing, Jian Su, Kevin Duh, and Xavier Carreras (Eds.). Association for Computational Linguistics, Austin, Texas, 2383\u20132392. Retrieved from https:\/\/aclanthology.org\/D16-1264\/"},{"key":"e_1_3_3_157_2","first-page":"3982","volume-title":"Proc. EMNLP-IJCNLP\u201919","author":"Reimers Nils","year":"2019","unstructured":"Nils Reimers and Iryna Gurevych. 2019. Sentence-BERT: Sentence embeddings using siamese BERT-networks. In Proc. EMNLP-IJCNLP\u201919. 3982\u20133992. Retrieved from https:\/\/aclanthology.org\/D19-1410"},{"key":"e_1_3_3_158_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-15712-8_32"},{"key":"e_1_3_3_159_2","doi-asserted-by":"publisher","DOI":"10.1145\/2939672.2939778"},{"key":"e_1_3_3_160_2","doi-asserted-by":"crossref","first-page":"842","DOI":"10.1162\/tacl_a_00349","article-title":"A primer in BERTology: What we know about how BERT works","volume":"8","author":"Rogers Anna","year":"2020","unstructured":"Anna Rogers, Olga Kovaleva, and Anna Rumshisky. 2020. A primer in BERTology: What we know about how BERT works. TACL 8 (2020), 842\u2013866. Retrieved from https:\/\/aclanthology.org\/2020.tacl-1.54","journal-title":"TACL"},{"key":"e_1_3_3_161_2","first-page":"767","volume-title":"Proc. NAACL-HLT","author":"Rothe Sascha","year":"2016","unstructured":"Sascha Rothe, Sebastian Ebert, and Hinrich Sch\u00fctze. 2016. Ultradense word embeddings by orthogonal transformation. In Proc. NAACL-HLT. 767\u2013777. Retrieved from https:\/\/aclanthology.org\/N16-1091"},{"key":"e_1_3_3_162_2","doi-asserted-by":"publisher","DOI":"10.1145\/3357384.3357859"},{"key":"e_1_3_3_163_2","volume-title":"Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers)","author":"Saad-Falcon Jon","year":"2024","unstructured":"Jon Saad-Falcon, Omar Khattab, Christopher Potts, and Matei Zaharia. 2024. ARES: An automated evaluation framework for retrieval-augmented generation systems. In Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers), Kevin Duh, Helena Gomez, and Steven Bethard (Eds.). Association for Computational Linguistics, Mexico City, Mexico. Retrieved from https:\/\/aclanthology.org\/2024.naacl-long.20\/"},{"key":"e_1_3_3_164_2","doi-asserted-by":"publisher","DOI":"10.1145\/3726302.3730343"},{"key":"e_1_3_3_165_2","unstructured":"Victor Sanh Lysandre Debut Julien Chaumond and Thomas Wolf. 2019. DistilBERT a distilled version of BERT: Smaller faster cheaper and lighter. (2019). arXiv:1910.01108. Retrieved from https:\/\/arxiv.org\/abs\/1910.01108"},{"key":"e_1_3_3_166_2","doi-asserted-by":"publisher","DOI":"10.1145\/3397271.3401286"},{"issue":"3","key":"e_1_3_3_167_2","first-page":"102925","article-title":"Learning interpretable word embeddings via bidirectional alignment of dimensions with semantic concepts","volume":"59","author":"Senel L\u00fctfi Kerem","year":"2022","unstructured":"L\u00fctfi Kerem Senel, Furkan Sahinuc, Veysel Y\u00fccesoy, Hinrich Sch\u00fctze, Tolga Cukur, and Aykut Koc. 2022. Learning interpretable word embeddings via bidirectional alignment of dimensions with semantic concepts. IPM 59, 3 (2022), 102925. Retrieved from https:\/\/www.sciencedirect.com\/science\/article\/pii\/S0306457322000498","journal-title":"IPM"},{"key":"e_1_3_3_168_2","doi-asserted-by":"publisher","DOI":"10.1017\/S1351324920000315"},{"key":"e_1_3_3_169_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/P19-1282"},{"key":"e_1_3_3_170_2","doi-asserted-by":"publisher","DOI":"10.1515\/9781400881970-018"},{"key":"e_1_3_3_171_2","volume-title":"Proc. ICLR\u201914, Workshop Track Proceedings","author":"Simonyan Karen","year":"2014","unstructured":"Karen Simonyan, Andrea Vedaldi, and Andrew Zisserman. 2014. Deep inside convolutional networks: Visualising image classification models and saliency maps. In Proc. ICLR\u201914, Workshop Track Proceedings. arXiv:1312.6034. Retrieved from http:\/\/arxiv.org\/abs\/1312.6034"},{"key":"e_1_3_3_172_2","volume-title":"Proc. ICLR\u201919","author":"Singh Chandan","year":"2019","unstructured":"Chandan Singh, W. James Murdoch, and Bin Yu. 2019. Hierarchical interpretations for neural network predictions. In Proc. ICLR\u201919. Retrieved from https:\/\/openreview.net\/forum?id=SkEqro0ctQ"},{"key":"e_1_3_3_173_2","doi-asserted-by":"publisher","DOI":"10.1145\/3289600.3290620"},{"key":"e_1_3_3_174_2","doi-asserted-by":"publisher","DOI":"10.1145\/3351095.3375234"},{"key":"e_1_3_3_175_2","volume-title":"Term Weighting Revisited","author":"Singhal Amit","year":"1997","unstructured":"Amit Singhal. 1997. Term Weighting Revisited. Ph. D. Dissertation. Cornell University. Retrieved from https:\/\/ecommons.cornell.edu\/server\/api\/core\/bitstreams\/ac2f078e-3307-454f-89dc-26693c44b4f3\/content"},{"key":"e_1_3_3_176_2","volume-title":"An Investigation of Dirichlet Prior Smoothing\u2019s Performance Advantage","author":"Smucker Mark D.","year":"2006","unstructured":"Mark D. Smucker and James Allan. 2006. An Investigation of Dirichlet Prior Smoothing\u2019s Performance Advantage. Technical Report IR-548. CIIR, U. Mass., Amherst. Retrieved from https:\/\/maroo.cs.umass.edu\/getpdf.php?id=694"},{"key":"e_1_3_3_177_2","first-page":"1631","volume-title":"Proc. EMNLP\u201913","author":"Socher Richard","year":"2013","unstructured":"Richard Socher, Alex Perelygin, Jean Wu, Jason Chuang, Christopher D. Manning, Andrew Ng, and Christopher Potts. 2013. Recursive deep models for semantic compositionality over a sentiment treebank. In Proc. EMNLP\u201913. 1631\u20131642. Retrieved from https:\/\/aclanthology.org\/D13-1170"},{"issue":"3","key":"e_1_3_3_178_2","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1007\/978-3-031-02180-0","article-title":"Explainable natural language processing","volume":"14","author":"S\u00f8gaard Anders","year":"2021","unstructured":"Anders S\u00f8gaard. 2021. Explainable natural language processing. Synthesis Lectures on Human Language Technologies 14, 3 (2021), 1\u2013123. Retrieved from https:\/\/link.springer.com\/book\/10.1007\/978-3-031-02180-0","journal-title":"Synthesis Lectures on Human Language Technologies"},{"key":"e_1_3_3_179_2","series-title":"NIPS \u201924","volume-title":"Proceedings of the 38th International Conference on Neural Information Processing Systems","author":"Su Zhaochen","year":"2025","unstructured":"Zhaochen Su, Jun Zhang, Xiaoye Qu, Tong Zhu, Yanshu Li, Jiashuo Sun, Juntao Li, Min Zhang, and Yu Cheng. 2025. CONFLICTBANK: A benchmark for evaluating knowledge conflicts in large language models. In Proceedings of the 38th International Conference on Neural Information Processing Systems (Vancouver, BC, Canada) (NIPS \u201924). Curran Associates Inc., Red Hook, NY, USA, Article 3280, 27 pages. Retrieved from https:\/\/openreview.net\/forum?id=wjHVmgBDzc"},{"key":"e_1_3_3_180_2","first-page":"9269","volume-title":"Proc. ICML\u201920","author":"Sundararajan Mukund","year":"2020","unstructured":"Mukund Sundararajan and Amir Najmi. 2020. The many shapley values for model explanation. In Proc. ICML\u201920, Vol. 119. 9269\u20139278. Retrieved from https:\/\/proceedings.mlr.press\/v119\/sundararajan20b.html"},{"key":"e_1_3_3_181_2","first-page":"3319","volume-title":"Proc. ICML\u201917","author":"Sundararajan Mukund","year":"2017","unstructured":"Mukund Sundararajan, Ankur Taly, and Qiqi Yan. 2017. Axiomatic attribution for deep networks. In Proc. ICML\u201917 (Sydney, NSW, Australia). 3319\u20133328. Retrieved from https:\/\/proceedings.mlr.press\/v70\/sundararajan17a.html"},{"key":"e_1_3_3_182_2","first-page":"719","volume-title":"Proc. of ACL\u201908: HLT","author":"Surdeanu Mihai","year":"2008","unstructured":"Mihai Surdeanu, Massimiliano Ciaramita, and Hugo Zaragoza. 2008. Learning to rank answers on large online QA collections. In Proc. of ACL\u201908: HLT. 719\u2013727. Retrieved from https:\/\/aclanthology.org\/P08-1082"},{"key":"e_1_3_3_183_2","doi-asserted-by":"publisher","DOI":"10.1145\/1277741.1277794"},{"key":"e_1_3_3_184_2","first-page":"4593","volume-title":"Proc. ACL\u201919","author":"Tenney Ian","year":"2019","unstructured":"Ian Tenney, Dipanjan Das, and Ellie Pavlick. 2019. BERT rediscovers the classical NLP pipeline. In Proc. ACL\u201919. 4593\u20134601. Retrieved from https:\/\/aclanthology.org\/P19-1452"},{"key":"e_1_3_3_185_2","volume-title":"Proc. ICLR\u201919","author":"Tenney Ian","year":"2019","unstructured":"Ian Tenney, Patrick Xia, Berlin Chen, Alex Wang, Adam Poliak, R. Thomas McCoy, Najoung Kim, Benjamin Van Durme, Samuel R. Bowman, Dipanjan Das, et\u00a0al. 2019. What do you learn from context? Probing for sentence structure in contextualized word representations. In Proc. ICLR\u201919. Retrieved from https:\/\/openreview.net\/forum?id=SJzSgnRcKX"},{"key":"e_1_3_3_186_2","volume-title":"Proc. NeurIPS\u201921 Datasets and Benchmarks Track (Round 2)","author":"Thakur Nandan","year":"2021","unstructured":"Nandan Thakur, Nils Reimers, Andreas R\u00fcckl\u00e9, Abhishek Srivastava, and Iryna Gurevych. 2021. BEIR: A heterogeneous benchmark for zero-shot evaluation of information retrieval models. In Proc. NeurIPS\u201921 Datasets and Benchmarks Track (Round 2). Retrieved from https:\/\/openreview.net\/forum?id=wCu6T5xFjeJ"},{"key":"e_1_3_3_187_2","first-page":"809","volume-title":"Proc. 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long Papers)","author":"Thorne James","year":"2018","unstructured":"James Thorne, Andreas Vlachos, Christos Christodoulopoulos, and Arpit Mittal. 2018. FEVER: A large-scale dataset for fact extraction and verification. In Proc. 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long Papers). 809\u2013819. DOI:10.18653\/v1\/N18-1074"},{"key":"e_1_3_3_188_2","first-page":"142","volume-title":"Proc. Natural Language Learning at HLT-NAACL\u201903","author":"Sang Erik F. Tjong Kim","year":"2003","unstructured":"Erik F. Tjong Kim Sang and Fien De Meulder. 2003. Introduction to the CoNLL-2003 shared task: Language-independent named entity recognition. In Proc. Natural Language Learning at HLT-NAACL\u201903. 142\u2013147. Retrieved from https:\/\/aclanthology.org\/W03-0419"},{"issue":"86","key":"e_1_3_3_189_2","first-page":"2579","article-title":"Visualizing data using t-SNE","volume":"9","author":"Maaten Laurens van der","year":"2008","unstructured":"Laurens van der Maaten and Geoffrey Hinton. 2008. Visualizing data using t-SNE. Journal of Machine Learning Research 9, 86 (2008), 2579\u20132605. Retrieved from http:\/\/jmlr.org\/papers\/v9\/vandermaaten08a.html","journal-title":"Journal of Machine Learning Research"},{"key":"e_1_3_3_190_2","volume-title":"Proc. NeurIPS\u201917","author":"Vaswani Ashish","year":"2017","unstructured":"Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, \u0141ukasz Kaiser, and Illia Polosukhin. 2017. Attention is All you Need. In Proc. NeurIPS\u201917, Vol. 30. Retrieved from https:\/\/proceedings.neurips.cc\/paper\/2017\/file\/3f5ee243547dee91fbd053c1c4a845aa-Paper.pdf"},{"key":"e_1_3_3_191_2","doi-asserted-by":"publisher","DOI":"10.1145\/3331184.3331377"},{"key":"e_1_3_3_192_2","first-page":"5797","volume-title":"Proc. ACL\u201919","author":"Voita Elena","year":"2019","unstructured":"Elena Voita, David Talbot, Fedor Moiseev, Rico Sennrich, and Ivan Titov. 2019. Analyzing multi-head self-attention: Specialized heads do the heavy lifting, the rest can be pruned. In Proc. ACL\u201919. Florence, Italy, 5797\u20135808. Retrieved from https:\/\/aclanthology.org\/P19-1580"},{"key":"e_1_3_3_193_2","doi-asserted-by":"publisher","DOI":"10.1145\/3471158.3472256"},{"key":"e_1_3_3_194_2","doi-asserted-by":"publisher","DOI":"10.1145\/3731120.3744592"},{"key":"e_1_3_3_195_2","first-page":"353","volume-title":"Proc. EMNLP\u201918 Workshop BlackboxNLP","author":"Wang Alex","year":"2018","unstructured":"Alex Wang, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel Bowman. 2018. GLUE: A multi-task benchmark and analysis platform for natural language understanding. In Proc. EMNLP\u201918 Workshop BlackboxNLP. 353\u2013355. Retrieved from https:\/\/aclanthology.org\/W18-5446"},{"key":"e_1_3_3_196_2","doi-asserted-by":"publisher","DOI":"10.1145\/1852102.1852106"},{"key":"e_1_3_3_197_2","first-page":"11","volume-title":"Proc. EMNLP-IJCNLP","author":"Wiegreffe Sarah","year":"2019","unstructured":"Sarah Wiegreffe and Yuval Pinter. 2019. Attention is not not explanation. In Proc. EMNLP-IJCNLP. 11\u201320. Retrieved from https:\/\/aclanthology.org\/D19-1002"},{"key":"e_1_3_3_198_2","doi-asserted-by":"publisher","DOI":"10.1145\/3576923"},{"key":"e_1_3_3_199_2","doi-asserted-by":"publisher","DOI":"10.1145\/3534928"},{"key":"e_1_3_3_200_2","volume-title":"Proc. ICLR\u201921","author":"Xiong Lee","year":"2021","unstructured":"Lee Xiong, Chenyan Xiong, Ye Li, Kwok-Fung Tang, Jialin Liu, Paul N. Bennett, Junaid Ahmed, and Arnold Overwijk. 2021. Approximate nearest neighbor negative contrastive learning for dense text retrieval. In Proc. ICLR\u201921. Retrieved from https:\/\/openreview.net\/forum?id=zeFrfgyZln"},{"key":"e_1_3_3_201_2","first-page":"8541","volume-title":"Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing","author":"Xu Rongwu","year":"2024","unstructured":"Rongwu Xu, Zehan Qi, Zhijiang Guo, Cunxiang Wang, Hongru Wang, Yue Zhang, and Wei Xu. 2024. Knowledge conflicts for LLMs: A survey. In Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing, Yaser Al-Onaizan, Mohit Bansal, and Yun-Nung Chen (Eds.). Association for Computational Linguistics, Miami, Florida, USA, 8541\u20138565. Retrieved from https:\/\/aclanthology.org\/2024.emnlp-main.486\/"},{"key":"e_1_3_3_202_2","volume-title":"Proceedings of the 10th ACM SIGIR\/14th International Conference on the Theory of Information Retrieval","author":"Xu Zhichao","year":"2024","unstructured":"Zhichao Xu, Hemank Lamba, Qingyao Ai, Joel R. Tetreault, and Alejandro Jaimes. 2024. CFE2: Counterfactual editing for search result explanation. In Proceedings of the 10th ACM SIGIR\/14th International Conference on the Theory of Information Retrieval. Retrieved from https:\/\/openreview.net\/forum?id=G3a15oOyJQ"},{"key":"e_1_3_3_203_2","doi-asserted-by":"publisher","DOI":"10.1145\/2983323.2983818"},{"key":"e_1_3_3_204_2","first-page":"2013","volume-title":"Proc. EMNLP\u201915","author":"Yang Yi","year":"2015","unstructured":"Yi Yang, Wen-tau Yih, and Christopher Meek. 2015. WikiQA: A challenge dataset for open-domain question answering. In Proc. EMNLP\u201915. 2013\u20132018. Retrieved from https:\/\/aclanthology.org\/D15-1237"},{"key":"e_1_3_3_205_2","doi-asserted-by":"crossref","first-page":"2369","DOI":"10.18653\/v1\/D18-1259","volume-title":"Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing","author":"Yang Zhilin","year":"2018","unstructured":"Zhilin Yang, Peng Qi, Saizheng Zhang, Yoshua Bengio, William Cohen, Ruslan Salakhutdinov, and Christopher D. Manning. 2018. HotpotQA: A dataset for diverse, explainable multi-hop question answering. In Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing, Ellen Riloff, David Chiang, Julia Hockenmaier, and Jun\u2019ichi Tsujii (Eds.). Association for Computational Linguistics, Brussels, Belgium, 2369\u20132380. Retrieved from https:\/\/aclanthology.org\/D18-1259\/"},{"key":"e_1_3_3_206_2","first-page":"1","volume-title":"Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies: Tutorials","author":"Yates Andrew","year":"2021","unstructured":"Andrew Yates, Rodrigo Nogueira, and Jimmy Lin. 2021. Pretrained transformers for text ranking: BERT and beyond. In Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies: Tutorials. 1\u20134. DOI:10.18653\/v1\/2021.naacl-tutorials.1"},{"key":"e_1_3_3_207_2","doi-asserted-by":"publisher","DOI":"10.1145\/1390334.1390435"},{"key":"e_1_3_3_208_2","doi-asserted-by":"publisher","DOI":"10.1145\/3477495.3532067"},{"key":"e_1_3_3_209_2","doi-asserted-by":"crossref","first-page":"3903","DOI":"10.18653\/v1\/2024.findings-acl.234","volume-title":"Findings of the Association for Computational Linguistics: ACL 2024","author":"Yuan Xiaowei","year":"2024","unstructured":"Xiaowei Yuan, Zhao Yang, Yequan Wang, Shengping Liu, Jun Zhao, and Kang Liu. 2024. Discerning and resolving knowledge conflicts through adaptive decoding with contextual information-entropy constraint. In Findings of the Association for Computational Linguistics: ACL 2024, Lun-Wei Ku, Andre Martins, and Vivek Srikumar (Eds.). Association for Computational Linguistics, Bangkok, Thailand, 3903\u20133922. Retrieved from https:\/\/aclanthology.org\/2024.findings-acl.234\/"},{"key":"e_1_3_3_210_2","volume-title":"Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing","author":"Yue Xiang","year":"2023","unstructured":"Xiang Yue, Boshi Wang, Ziru Chen, Kai Zhang, Yu Su, and Huan Sun. 2023. Automatic evaluation of attribution by large language models. In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing. Retrieved from https:\/\/openreview.net\/forum?id=jVa7tFQw9N"},{"key":"e_1_3_3_211_2","series-title":"WWW \u201920","doi-asserted-by":"crossref","first-page":"418","DOI":"10.1145\/3366423.3380126","volume-title":"Proceedings of The Web Conference 2020","author":"Zamani Hamed","year":"2020","unstructured":"Hamed Zamani, Susan Dumais, Nick Craswell, Paul Bennett, and Gord Lueck. 2020. Generating clarifying questions for information retrieval. In Proceedings of The Web Conference 2020 (Taipei, Taiwan) (WWW \u201920). Association for Computing Machinery, New York, NY, USA, 418\u2013428. DOI:10.1145\/3366423.3380126"},{"key":"e_1_3_3_212_2","doi-asserted-by":"publisher","DOI":"10.1145\/3340531.3412772"},{"key":"e_1_3_3_213_2","doi-asserted-by":"publisher","DOI":"10.1561\/1500000081"},{"key":"e_1_3_3_214_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-10590-1_53"},{"key":"e_1_3_3_215_2","doi-asserted-by":"publisher","DOI":"10.1145\/984321.984322"},{"key":"e_1_3_3_216_2","doi-asserted-by":"publisher","DOI":"10.1145\/3397271.3401325"},{"key":"e_1_3_3_217_2","doi-asserted-by":"publisher","DOI":"10.1145\/3331184.3331208"},{"key":"e_1_3_3_218_2","unstructured":"Sheng Zhang Xiaodong Liu Jingjing Liu Jianfeng Gao Kevin Duh and Benjamin Van Durme. 2018. ReCoRD: Bridging the gap between human and machine commonsense reading comprehension. (2018). arXiv:1810.12885. Retrieved from http:\/\/arxiv.org\/abs\/1810.12885"},{"key":"e_1_3_3_219_2","volume-title":"Proc. ICLR\u201920","author":"Zhang Tianyi","year":"2020","unstructured":"Tianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger, and Yoav Artzi. 2020. BERTScore: Evaluating text generation with BERT. In Proc. ICLR\u201920. Retrieved from https:\/\/openreview.net\/forum?id=SkeHuCVFDr"},{"key":"e_1_3_3_220_2","doi-asserted-by":"publisher","DOI":"10.1145\/3437963.3441796"},{"key":"e_1_3_3_221_2","doi-asserted-by":"publisher","DOI":"10.1145\/3529755"}],"container-title":["ACM Computing Surveys"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3801957","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,5,1]],"date-time":"2026-05-01T10:12:33Z","timestamp":1777630353000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3801957"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,5,1]]},"references-count":220,"journal-issue":{"issue":"11","published-print":{"date-parts":[[2026,8,30]]}},"alternative-id":["10.1145\/3801957"],"URL":"https:\/\/doi.org\/10.1145\/3801957","relation":{},"ISSN":["0360-0300","1557-7341"],"issn-type":[{"value":"0360-0300","type":"print"},{"value":"1557-7341","type":"electronic"}],"subject":[],"published":{"date-parts":[[2026,5,1]]},"assertion":[{"value":"2022-12-05","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2026-02-10","order":2,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2026-05-01","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}