{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,5,13]],"date-time":"2025-05-13T22:00:09Z","timestamp":1747173609929,"version":"3.40.5"},"reference-count":51,"publisher":"Cambridge University Press (CUP)","issue":"3","license":[{"start":{"date-parts":[[2020,3,6]],"date-time":"2020-03-06T00:00:00Z","timestamp":1583452800000},"content-version":"unspecified","delay-in-days":0,"URL":"https:\/\/www.cambridge.org\/core\/terms"}],"content-domain":{"domain":["cambridge.org"],"crossmark-restriction":true},"short-container-title":["Nat. Lang. Eng."],"published-print":{"date-parts":[[2021,5]]},"abstract":"<jats:title>Abstract<\/jats:title><jats:p>Technical writing in professional environments, such as user manual authoring, requires the use of uniform language. Nonuniform language refers to sentences in a technical document that are intended to have the same meaning within a similar context, but use different words or writing style. Addressing this nonuniformity problem requires the performance of two tasks. The first task, which we named nonuniform language detection (NLD), is detecting such sentences. We propose an NLD method that utilizes different similarity algorithms at lexical, syntactic, semantic and pragmatic levels. Different features are extracted and integrated by applying a machine learning classification method. The second task, which we named nonuniform language correction (NLC), is deciding which sentence among the detected ones is more appropriate for that context. To address this problem, we propose an NLC method that combines contraction removal, near-synonym choice, and text readability comparison. We tested our methods using smartphone user manuals. We finally compared our methods against state-of-the-art methods in paraphrase detection (for NLD) and against expert annotators (for both NLD and NLC). The experiments demonstrate that the proposed methods achieve performance that matches expert annotators.<\/jats:p>","DOI":"10.1017\/s1351324920000133","type":"journal-article","created":{"date-parts":[[2020,3,6]],"date-time":"2020-03-06T05:58:36Z","timestamp":1583474316000},"page":"293-314","update-policy":"https:\/\/doi.org\/10.1017\/policypage","source":"Crossref","is-referenced-by-count":1,"title":["Nonuniform language in technical writing: Detection and correction"],"prefix":"10.1017","volume":"27","author":[{"given":"Weibo","family":"Wang","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Aminul","family":"Islam","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Abidalrahman","family":"Moh\u2019d","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-9021-7566","authenticated-orcid":false,"given":"Axel J.","family":"Soto","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Evangelos E.","family":"Milios","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"56","published-online":{"date-parts":[[2020,3,6]]},"reference":[{"key":"S1351324920000133_ref46","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-41230-1_24"},{"key":"S1351324920000133_ref35","doi-asserted-by":"publisher","DOI":"10.1145\/1242572.1242592"},{"key":"S1351324920000133_ref50","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-39357-0_17"},{"key":"S1351324920000133_ref51","unstructured":"Wu, Z. and Palmer, M. (1994). Verbs Semantics and Lexical Selection. Available at: Demo URL: http:\/\/ws4jdemo.appspot.com\/?mode=w&s1=&w1=photo&s2=&w2=video (Accessed 01 December 2015)."},{"key":"S1351324920000133_ref27","doi-asserted-by":"publisher","DOI":"10.1145\/1376815.1376819"},{"key":"S1351324920000133_ref21","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-540-68125-0_87"},{"key":"S1351324920000133_ref11","doi-asserted-by":"publisher","DOI":"10.1037\/h0076540"},{"key":"S1351324920000133_ref4","doi-asserted-by":"publisher","DOI":"10.1145\/3110025.3122119"},{"key":"S1351324920000133_ref31","unstructured":"Ke\u0161elj, V. and Cercone, N. (2004). CNG method with weighted voting. In Ad-hoc Authorship Attribution Competition. Proceedings 2004 Joint International Conference of the Association for Literary and Linguistic Computing and the Association for Computers and the Humanities (ALLC\/ACH 2004)."},{"key":"S1351324920000133_ref29","first-page":"41","article-title":"Near-synonym choice using a 5-gram language model","volume":"46","author":"Islam","year":"2010","journal-title":"Research in Computing Sciences"},{"key":"S1351324920000133_ref40","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/W16-1617"},{"key":"S1351324920000133_ref2","doi-asserted-by":"publisher","DOI":"10.1613\/jair.2985"},{"key":"S1351324920000133_ref6","unstructured":"Brants, T. and Franz, A. (2009). Web 1T 5-gram, 10 European languages version 1. LDC2009T25. Web Download. Philadelphia: Linguistic Data Consortium."},{"key":"S1351324920000133_ref20","unstructured":"Gionis, A. , Indyk, P. and Motwani, R. (1999). Similarity search in high dimensions via hashing. In Proceedings of the 25th International Conference on Very Large Data Bases, VLDB \u201999, San Francisco, CA, USA: Morgan Kaufmann Publishers Inc., pp. 518\u2013529"},{"key":"S1351324920000133_ref34","unstructured":"LG (2009). LG600G User Guide. Available at: https:\/\/www.manualslib.com\/manual\/92956\/Lg-Lg600g.html#product-LG600G (Accessed 15 December 2015)."},{"key":"S1351324920000133_ref24","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-32612-7_10"},{"key":"S1351324920000133_ref1","doi-asserted-by":"publisher","DOI":"10.1016\/j.ipm.2018.06.005"},{"key":"S1351324920000133_ref12","doi-asserted-by":"publisher","DOI":"10.1111\/j.1467-9922.2009.00508.x"},{"key":"S1351324920000133_ref39","doi-asserted-by":"crossref","unstructured":"Mueller, J. and Thyagarajan, A. (2016). Siamese recurrent architectures for learning sentence similarity. In Proceedings of the Thirtieth AAAI Conference on Artificial Intelligence, AAAI \u201916. AAAI Press, pp. 2786\u20132792.","DOI":"10.1609\/aaai.v30i1.10350"},{"key":"S1351324920000133_ref25","unstructured":"Inkpen, D.Z. (2007). Near-synonym choice in an intelligent thesaurus. In HLT-NAACL, Rochester, NY, April 22\u201327, 2007, pp. 356\u2013363."},{"key":"S1351324920000133_ref15","doi-asserted-by":"publisher","DOI":"10.1162\/COLI_a_00255"},{"key":"S1351324920000133_ref33","first-page":"37","article-title":"Metonymy: Developing a cognitive linguistic view","volume":"9","author":"K\u00f6vecses","year":"1998","journal-title":"Cognitive Linguistics (Includes Cognitive Linguistic Bibliography)"},{"key":"S1351324920000133_ref42","unstructured":"Samsung (2011). Samsung 010505d5 cell phone user manual. Available at: http:\/\/cellphone.manualsonline.com\/manuals\/mfg\/samsung\/010505d5.html?p=53 (Accessed 01 December 2015)."},{"key":"S1351324920000133_ref17","doi-asserted-by":"publisher","DOI":"10.2190\/T6EM-UTT0-EL6J-59N9"},{"key":"S1351324920000133_ref7","doi-asserted-by":"crossref","unstructured":"Chen, Q. , Hu, Q. , Huang, J. X. and He, L. (2018). CA-RNN: Using context-aligned recurrent neural networks for modeling sentence similarity. In Thirty-Second AAAI Conference on Artificial Intelligence, AAAI Press, pp. 265\u2013273.","DOI":"10.1609\/aaai.v32i1.11273"},{"key":"S1351324920000133_ref45","doi-asserted-by":"publisher","DOI":"10.1145\/2682571.2797068"},{"key":"S1351324920000133_ref3","unstructured":"Apple Inc. (2015). iPhone User Guide For iOS 8.4 Software. Available at: https:\/\/manuals.info.apple.com\/MANUALS\/1000\/MA1565\/en_US\/iphone_user_guide.pdf (Accessed 01 December 2015)."},{"volume-title":"Natural Language Processing with Python","year":"2009","author":"Bird","key":"S1351324920000133_ref5"},{"key":"S1351324920000133_ref37","doi-asserted-by":"crossref","unstructured":"Mei, J. , Kou, X. , Yao, Z. , Rau-Chaplin, A. , Islam, A. , Moh\u2019d, A. and Milios, E.E. (2015). Efficient Computation of Co-Occurrence Based Word Relatedness. Available at Demo URL: http:\/\/ares.research.cs.dal.ca\/gtm\/ (Accessed 01 December 2015).","DOI":"10.1145\/2682571.2797088"},{"key":"S1351324920000133_ref8","first-page":"463","article-title":"A fast algorithm for computing longest common subsequences of small alphabet size","volume":"13","author":"Chin","year":"1991","journal-title":"Journal of Information Processing"},{"key":"S1351324920000133_ref28","doi-asserted-by":"crossref","unstructured":"Islam, A. and Inkpen, D. (2009). Real-word spelling correction using google web 1t n-gram data set. In Proceedings of the 18th ACM Conference on Information and Knowledge Management. ACM, pp. 1689\u20131692.","DOI":"10.1145\/1645953.1646205"},{"key":"S1351324920000133_ref41","doi-asserted-by":"publisher","DOI":"10.3115\/1621969.1621980"},{"key":"S1351324920000133_ref18","unstructured":"Feng, L. , Jansche, M. , Huenerfauth, M. and Elhadad, N. (2010). A comparison of features for automatic readability assessment. In Proceedings of the 23rd International Conference on Computational Linguistics: Posters. Association for Computational Linguistics, pp. 276\u2013284."},{"key":"S1351324920000133_ref13","first-page":"84","article-title":"Text readability and intuitive simplification: A comparison of readability formulas","volume":"23","author":"Crossley","year":"2011","journal-title":"Reading in a Foreign Language"},{"key":"S1351324920000133_ref36","doi-asserted-by":"publisher","DOI":"10.1017\/CBO9780511894664"},{"key":"S1351324920000133_ref49","doi-asserted-by":"publisher","DOI":"10.1109\/IJCNN.2018.8489222"},{"key":"S1351324920000133_ref23","doi-asserted-by":"publisher","DOI":"10.1177\/002194366900600202"},{"key":"S1351324920000133_ref14","doi-asserted-by":"publisher","DOI":"10.3115\/1687878.1687944"},{"key":"S1351324920000133_ref9","doi-asserted-by":"publisher","DOI":"10.1037\/h0026256"},{"key":"S1351324920000133_ref44","unstructured":"Socher, R. , Huang, E.H. , Pennington, J. , Ng, A.Y. and Manning, C.D. (2011). Dynamic pooling and unfolding recursive autoencoders for paraphrase detection. In Proceedings of the 24th International Conference on Neural Information Processing Systems, NIPS \u201911. Red Hook, NY, USA: Curran Associates Inc., pp. 801\u2013809."},{"key":"S1351324920000133_ref43","first-page":"3","article-title":"Automated readability index","author":"Senter","year":"1967","journal-title":"Wright-Patterson Air Force Base. AMRL-TR-6620"},{"key":"S1351324920000133_ref32","doi-asserted-by":"publisher","DOI":"10.21236\/ADA006655"},{"key":"S1351324920000133_ref16","unstructured":"Devlin, J. , Chang, M.-W. , Lee, K. and Toutanova, K. (2019). BERT: Pre-training of deep bidirectional transformers for language understanding. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics, Minneapolis, Minnesota: Association for Computational Linguistics, pp. 4171\u20134186."},{"key":"S1351324920000133_ref48","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/D16-1194"},{"key":"S1351324920000133_ref22","doi-asserted-by":"publisher","DOI":"10.3758\/BF03195564"},{"key":"S1351324920000133_ref10","first-page":"35","article-title":"How evaluation guides AI research: The message still counts more than the medium","volume":"9","author":"Cohen","year":"1988","journal-title":"AI Magazine"},{"key":"S1351324920000133_ref19","doi-asserted-by":"publisher","DOI":"10.1177\/001316447303300309"},{"key":"S1351324920000133_ref38","doi-asserted-by":"publisher","DOI":"10.1093\/ijl\/3.4.235"},{"volume-title":"The Nature of Statistical Learning Theory","year":"2013","author":"Vapnik","key":"S1351324920000133_ref47"},{"key":"S1351324920000133_ref30","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-30353-1_29"},{"key":"S1351324920000133_ref26","doi-asserted-by":"publisher","DOI":"10.1007\/3-540-56024-6_18"}],"container-title":["Natural Language Engineering"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.cambridge.org\/core\/services\/aop-cambridge-core\/content\/view\/S1351324920000133","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2022,10,17]],"date-time":"2022-10-17T20:12:45Z","timestamp":1666037565000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.cambridge.org\/core\/product\/identifier\/S1351324920000133\/type\/journal_article"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,3,6]]},"references-count":51,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2021,5]]}},"alternative-id":["S1351324920000133"],"URL":"https:\/\/doi.org\/10.1017\/s1351324920000133","relation":{},"ISSN":["1351-3249","1469-8110"],"issn-type":[{"type":"print","value":"1351-3249"},{"type":"electronic","value":"1469-8110"}],"subject":[],"published":{"date-parts":[[2020,3,6]]},"assertion":[{"value":"\u00a9 Cambridge University Press 2020","name":"copyright","label":"Copyright","group":{"name":"copyright_and_licensing","label":"Copyright and Licensing"}}]}}