{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,19]],"date-time":"2026-05-19T07:17:32Z","timestamp":1779175052768,"version":"3.51.4"},"publisher-location":"New York, NY, USA","reference-count":24,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,10,17]],"date-time":"2022-10-17T00:00:00Z","timestamp":1665964800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,10,17]]},"DOI":"10.1145\/3511808.3557673","type":"proceedings-article","created":{"date-parts":[[2022,10,16]],"date-time":"2022-10-16T01:22:22Z","timestamp":1665883342000},"page":"3786-3790","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":14,"title":["Probing the Robustness of Pre-trained Language Models for Entity Matching"],"prefix":"10.1145","author":[{"given":"Mehdi","family":"Akbarian Rastaghi","sequence":"first","affiliation":[{"name":"University of Alberta, Edmonton, AB, Canada"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Ehsan","family":"Kamalloo","sequence":"additional","affiliation":[{"name":"University of Alberta, Edmonton, AB, Canada"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Davood","family":"Rafiei","sequence":"additional","affiliation":[{"name":"University of Alberta, Edmonton, AB, Canada"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2022,10,17]]},"reference":[{"key":"e_1_3_2_1_1_1","volume-title":"Neural Networks for Entity Matching: A Survey. ACM Transactions on Knowledge Discovery from Data","author":"Barlaug Nils","year":"2021","unstructured":"Nils Barlaug and Jon Atle Gulla . 2021. Neural Networks for Entity Matching: A Survey. ACM Transactions on Knowledge Discovery from Data ( 2021 ). https:\/\/doi.org\/10.1145\/3442200 10.1145\/3442200 Nils Barlaug and Jon Atle Gulla. 2021. Neural Networks for Entity Matching: A Survey. ACM Transactions on Knowledge Discovery from Data (2021). https:\/\/doi.org\/10.1145\/3442200"},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1145\/3442188.3445922"},{"key":"e_1_3_2_1_3_1","volume":"200","author":"Bilenko Mikhail","unstructured":"Mikhail Bilenko and Raymond J. Mooney. 200 3. Adaptive Duplicate Detection Using Learnable String Similarity Measures. Proceedings of the Ninth ACM SIGKDD International Conference on Knowledge Discovery and Data Mining. https:\/\/doi.org\/10.1145\/956750.956759 10.1145\/956750.956759 Mikhail Bilenko and Raymond J. Mooney. 2003. Adaptive Duplicate Detection Using Learnable String Similarity Measures. Proceedings of the Ninth ACM SIGKDD International Conference on Knowledge Discovery and Data Mining. https:\/\/doi.org\/10.1145\/956750.956759","journal-title":"Raymond J. Mooney."},{"key":"e_1_3_2_1_4_1","volume-title":"Advances in Neural Information Processing Systems","author":"Brown Tom B.","year":"2020","unstructured":"Tom B. Brown , Benjamin Mann , Nick Ryder , Melanie Subbiah , Jared Kaplan , Prafulla Dhariwal , Arvind Neelakantan , Pranav Shyam , Girish Sastry , Amanda Askell , Sandhini Agarwal , Ariel Herbert-Voss , Gretchen Krueger , Tom Henighan , Rewon Child , Aditya Ramesh , Daniel M. Ziegler , Jeffrey Wu , Clemens Winter , Christopher Hesse , Mark Chen , Eric Sigler , Mateusz Litwin , Scott Gray , Benjamin Chess , Jack Clark , Christopher Berner , Sam McCandlish , Alec Radford , Ilya Sutskever , and Dario Amodei . 2020 . Language models are few-shot learners . Advances in Neural Information Processing Systems , Vol. 2020-December (2020). https:\/\/proceedings.neurips.cc\/paper\/2020\/file\/1457c0d6bfcb4967418bfb8ac142f64a-Paper.pdf Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020. Language models are few-shot learners. Advances in Neural Information Processing Systems, Vol. 2020-December (2020). https:\/\/proceedings.neurips.cc\/paper\/2020\/file\/1457c0d6bfcb4967418bfb8ac142f64a-Paper.pdf"},{"key":"e_1_3_2_1_5_1","volume-title":"NAACL HLT 2019 - 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies - Proceedings of the Conference (2019","author":"Devlin Jacob","year":"2019","unstructured":"Jacob Devlin , Ming Wei Chang , Kenton Lee , and Kristina Toutanova . 2019 . BERT: Pre-training of deep bidirectional transformers for language understanding . NAACL HLT 2019 - 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies - Proceedings of the Conference (2019 ). https:\/\/doi.org\/10.18653\/v1\/N19--1423 10.18653\/v1 Jacob Devlin, Ming Wei Chang, Kenton Lee, and Kristina Toutanova. 2019. BERT: Pre-training of deep bidirectional transformers for language understanding. NAACL HLT 2019 - 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies - Proceedings of the Conference (2019). https:\/\/doi.org\/10.18653\/v1\/N19--1423"},{"key":"e_1_3_2_1_6_1","volume-title":"Verykios","author":"Elmagarmid Ahmed K.","year":"2007","unstructured":"Ahmed K. Elmagarmid , Panagiotis G. Ipeirotis , and Vassilios S . Verykios . 2007 . Duplicate Record Detection: A Survey. IEEE Trans. on Knowl. and Data Eng . (2007). https:\/\/doi.org\/10.1109\/TKDE.1990.10000 10.1109\/TKDE.1990.10000 Ahmed K. Elmagarmid, Panagiotis G. Ipeirotis, and Vassilios S. Verykios. 2007. Duplicate Record Detection: A Survey. IEEE Trans. on Knowl. and Data Eng. (2007). https:\/\/doi.org\/10.1109\/TKDE.1990.10000"},{"key":"e_1_3_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.14778\/1687627.1687674"},{"key":"e_1_3_2_1_8_1","volume-title":"ACL 2019 - 57th Annual Meeting of the Association for Computational Linguistics, Proceedings of the Conference (2020","author":"Kasai Jungo","year":"2020","unstructured":"Jungo Kasai , Kun Qian , Sairam Gurajada , Yunyao Li , and Lucian Popa . 2020 . Low-resource deep entity resolution with transfer and active learning . ACL 2019 - 57th Annual Meeting of the Association for Computational Linguistics, Proceedings of the Conference (2020 ). https:\/\/doi.org\/10.18653\/v1\/p19--1586 10.18653\/v1 Jungo Kasai, Kun Qian, Sairam Gurajada, Yunyao Li, and Lucian Popa. 2020. Low-resource deep entity resolution with transfer and active learning. ACL 2019 - 57th Annual Meeting of the Association for Computational Linguistics, Proceedings of the Conference (2020). https:\/\/doi.org\/10.18653\/v1\/p19--1586"},{"key":"e_1_3_2_1_9_1","volume-title":"Magellan: Toward building entity matching management systems","author":"Konda Pradap Venkatramanan","year":"2018","unstructured":"Pradap Venkatramanan Konda . 2018 . Magellan: Toward building entity matching management systems . http:\/\/www.vldb.org\/pvldb\/vol9\/p1197-pkonda.pdf Pradap Venkatramanan Konda. 2018. Magellan: Toward building entity matching management systems. http:\/\/www.vldb.org\/pvldb\/vol9\/p1197-pkonda.pdf"},{"key":"e_1_3_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.14778\/1920841.1920904"},{"key":"e_1_3_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2020.acl-main.45"},{"key":"e_1_3_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1145\/3431816"},{"key":"e_1_3_2_1_13_1","volume-title":"Roberta: A robustly optimized bert pretraining approach. arXiv preprint arXiv:1907.11692","author":"Liu Yinhan","year":"2019","unstructured":"Yinhan Liu , Myle Ott , Naman Goyal , Jingfei Du , Mandar Joshi , Danqi Chen , Omer Levy , Mike Lewis , Luke Zettlemoyer , and Veselin Stoyanov . 2019 . Roberta: A robustly optimized bert pretraining approach. arXiv preprint arXiv:1907.11692 (2019). https:\/\/doi.org\/10.48550\/arXiv.1907.11692 10.48550\/arXiv.1907.11692 Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019. Roberta: A robustly optimized bert pretraining approach. arXiv preprint arXiv:1907.11692 (2019). https:\/\/doi.org\/10.48550\/arXiv.1907.11692"},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/P19-1334"},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1145\/3183713.3196926"},{"key":"e_1_3_2_1_16_1","volume-title":"ACL 2019 - 57th Annual Meeting of the Association for Computational Linguistics, Proceedings of the Conference. https:\/\/doi.org\/10","author":"Niven Timothy","year":"2020","unstructured":"Timothy Niven and Hung Yu Kao . 2020 . Probing neural network comprehension of natural language arguments . ACL 2019 - 57th Annual Meeting of the Association for Computational Linguistics, Proceedings of the Conference. https:\/\/doi.org\/10 .18653\/v1\/p19--1459 10.18653\/v1 Timothy Niven and Hung Yu Kao. 2020. Probing neural network comprehension of natural language arguments. ACL 2019 - 57th Annual Meeting of the Association for Computational Linguistics, Proceedings of the Conference. https:\/\/doi.org\/10.18653\/v1\/p19--1459"},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.14778\/3467861.3467878"},{"key":"e_1_3_2_1_18_1","volume-title":"Improving Language Understanding by Generative Pre-Training. OpenAI blog","author":"Radford Alec","year":"2018","unstructured":"Alec Radford , Karthik Narasimhan , Tim Salimans , and Ilya Sutskever . 2018. Improving Language Understanding by Generative Pre-Training. OpenAI blog ( 2018 ). https:\/\/www.gwern.net\/docs\/www\/s3-us-west-2.amazonaws.com\/d73fdc5ffa8627bce44dcda2fc012da638ffb158.pdf Alec Radford, Karthik Narasimhan, Tim Salimans, and Ilya Sutskever. 2018. Improving Language Understanding by Generative Pre-Training. OpenAI blog (2018). https:\/\/www.gwern.net\/docs\/www\/s3-us-west-2.amazonaws.com\/d73fdc5ffa8627bce44dcda2fc012da638ffb158.pdf"},{"key":"e_1_3_2_1_19_1","volume-title":"David Luan Rewon Child, and Dario Amodei","author":"Alec Radford Ilya Sutskever","year":"2019","unstructured":"Ilya Sutskever Alec Radford , Jeffrey Wu , David Luan Rewon Child, and Dario Amodei . 2019 . Language Models are Unsupervised Multitask Learners. OpenAI Blog ( 2019). https:\/\/d4mucfpksywv.cloudfront.net\/better-language-models\/language-models.pdf Ilya Sutskever Alec Radford, Jeffrey Wu, David Luan Rewon Child, and Dario Amodei. 2019. Language Models are Unsupervised Multitask Learners. OpenAI Blog (2019). https:\/\/d4mucfpksywv.cloudfront.net\/better-language-models\/language-models.pdf"},{"key":"e_1_3_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.1145\/3035918.3058739"},{"key":"e_1_3_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1145\/3488560.3498486"},{"key":"e_1_3_2_1_22_1","volume-title":"Transformers: State-of-the-Art Natural Language Processing. Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing: System Demonstrations. https:\/\/doi.org\/10.1865","author":"Wolf Thomas","year":"2020","unstructured":"Thomas Wolf , Lysandre Debut , Victor Sanh , Julien Chaumond , Clement Delangue , Anthony Moi , Pierric Cistac , Tim Rault , R\u00e9mi Louf , Morgan Funtowicz , Joe Davison , Sam Shleifer , Patrick von Platen , Clara Ma , Yacine Jernite , Julien Plu , Canwen Xu , Teven Le Scao , Sylvain Gugger , Mariama Drame , Quentin Lhoest , and Alexander M. Rush . 2020 . Transformers: State-of-the-Art Natural Language Processing. Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing: System Demonstrations. https:\/\/doi.org\/10.1865 3\/v1\/ 2020 .emnlp-demos.6 10.18653\/v1 Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, R\u00e9mi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander M. Rush. 2020. Transformers: State-of-the-Art Natural Language Processing. Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing: System Demonstrations. https:\/\/doi.org\/10.18653\/v1\/2020.emnlp-demos.6"},{"key":"e_1_3_2_1_23_1","volume-title":"Auto-EM: End-to-End Fuzzy Entity-Matching Using Pre-Trained Deep Models and Transfer Learning. The World Wide Web Conference. https:\/\/doi.org\/10","author":"Zhao Chen","year":"2019","unstructured":"Chen Zhao and Yeye He . 2019 . Auto-EM: End-to-End Fuzzy Entity-Matching Using Pre-Trained Deep Models and Transfer Learning. The World Wide Web Conference. https:\/\/doi.org\/10 .1145\/3308558.3313578 10.1145\/3308558.3313578 Chen Zhao and Yeye He. 2019. Auto-EM: End-to-End Fuzzy Entity-Matching Using Pre-Trained Deep Models and Transfer Learning. The World Wide Web Conference. https:\/\/doi.org\/10.1145\/3308558.3313578"},{"key":"e_1_3_2_1_24_1","volume-title":"Application of Weighted Cross-Entropy Loss Function in Intrusion Detection. Journal of Computer and Communications","author":"Zhou Ziyun","year":"2021","unstructured":"Ziyun Zhou , Hong Huang , and Binhao Fang . 2021. Application of Weighted Cross-Entropy Loss Function in Intrusion Detection. Journal of Computer and Communications ( 2021 ). https:\/\/doi.org\/10.4236\/jcc.2021.911001 10.4236\/jcc.2021.911001 Ziyun Zhou, Hong Huang, and Binhao Fang. 2021. Application of Weighted Cross-Entropy Loss Function in Intrusion Detection. Journal of Computer and Communications (2021). https:\/\/doi.org\/10.4236\/jcc.2021.911001"}],"event":{"name":"CIKM '22: The 31st ACM International Conference on Information and Knowledge Management","location":"Atlanta GA USA","acronym":"CIKM '22","sponsor":["SIGWEB ACM Special Interest Group on Hypertext, Hypermedia, and Web","SIGIR ACM Special Interest Group on Information Retrieval"]},"container-title":["Proceedings of the 31st ACM International Conference on Information &amp; Knowledge Management"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3511808.3557673","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3511808.3557673","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T17:48:49Z","timestamp":1750182529000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3511808.3557673"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,10,17]]},"references-count":24,"alternative-id":["10.1145\/3511808.3557673","10.1145\/3511808"],"URL":"https:\/\/doi.org\/10.1145\/3511808.3557673","relation":{},"subject":[],"published":{"date-parts":[[2022,10,17]]},"assertion":[{"value":"2022-10-17","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}