{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,19]],"date-time":"2025-06-19T04:14:07Z","timestamp":1750306447206,"version":"3.41.0"},"reference-count":18,"publisher":"Association for Computing Machinery (ACM)","issue":"1","license":[{"start":{"date-parts":[[2015,9,29]],"date-time":"2015-09-29T00:00:00Z","timestamp":1443484800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["SIGKDD Explor. Newsl."],"published-print":{"date-parts":[[2015,9,29]]},"abstract":"<jats:p>Much of the English in text documents today comes from nonnative speakers. Web searches are also conducted very often by non-native speakers. Though highly qualified in their respective fields, these speakers could potentially make errors in collocation, e.g., \"dark money\" and \"stock agora\" (instead of the more appropriate English expressions \"black money\" and \"stock market\" respectively). These may arise due to literal translation from the respective speaker's native language or other factors. Such errors could cause problems in contexts such as querying over Web pages, correct understanding of text documents and more. This paper proposes a framework called CollOrder to detect such collocation errors and suggest correctly ordered collocated responses for improving the semantics. This framework integrates machine learning approaches with natural language processing techniques, proposing suitable heuristics to provide responses to collocation errors, ranked in the order of correctness. We discuss the proposed framework with algorithms and experimental evaluation in this paper. We claim that it would be useful in semantically enhancing Web querying e.g., financial news, online shopping etc. It would also help in providing automated error correction in machine translated documents and offering assistance to people using ESL tools.<\/jats:p>","DOI":"10.1145\/2830544.2830548","type":"journal-article","created":{"date-parts":[[2015,9,29]],"date-time":"2015-09-29T19:22:29Z","timestamp":1443554549000},"page":"14-23","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":8,"title":["A Framework for Collocation Error Correction in Web Pages and Text Documents"],"prefix":"10.1145","volume":"17","author":[{"given":"Alan","family":"Varghese","sequence":"first","affiliation":[{"name":"Montclair State University, Montclair, NJ, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Aparna S.","family":"Varde","sequence":"additional","affiliation":[{"name":"Montclair State University, Montclair, NJ, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jing","family":"Peng","sequence":"additional","affiliation":[{"name":"Montclair State University, Montclair, NJ, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Eileen","family":"Fitzpatrick","sequence":"additional","affiliation":[{"name":"Montclair State University, Montclair, NJ, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2015,9,29]]},"reference":[{"key":"e_1_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1016\/B978-1-55860-377-6.50023-2"},{"key":"e_1_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1145\/1498759.1498815"},{"key":"e_1_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.5555\/2145432.2145445"},{"key":"e_1_2_1_4_1","first-page":"43","volume-title":"Fitzpatrick","author":"Deane P.","year":"2007","unstructured":"Deane , P. and Higgins , D. , Using Singular-Value Decomposition on local word contexts to derive a measure of constructional similarity , in Fitzpatrick , E. (editor). Corpus Linguistics Beyond the Word : Corpus Research from Phrase to Discourse, 2007 , Rodopi , pp. 43 -- 58 . Deane, P. and Higgins, D., Using Singular-Value Decomposition on local word contexts to derive a measure of constructional similarity, in Fitzpatrick, E. (editor). Corpus Linguistics Beyond the Word: Corpus Research from Phrase to Discourse, 2007, Rodopi, pp. 43--58."},{"key":"e_1_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1080\/09588220802343561"},{"key":"e_1_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1145\/1807167.1807228"},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1145\/1951365.1951442"},{"key":"e_1_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.3115\/1075096.1075150"},{"key":"e_1_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.5555\/1609843.1609850"},{"key":"e_1_2_1_10_1","first-page":"449","article-title":"Generating Typed Dependency Parses from Phrase Structure Parses","volume":"6","author":"Marneffe M.","year":"2006","unstructured":"Marneffe , M. , MacCartney , B. and Manning , C ., Generating Typed Dependency Parses from Phrase Structure Parses . LREC , 2006 , ( 6 ): 449 -- 454 . Marneffe, M., MacCartney, B. and Manning, C., Generating Typed Dependency Parses from Phrase Structure Parses. LREC, 2006, (6):449--454.","journal-title":"LREC"},{"key":"e_1_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1093\/ijl\/3.4.235"},{"key":"e_1_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1145\/1449715.1449736"},{"key":"e_1_2_1_13_1","first-page":"259","article-title":"Automatic Classification of Article Errors in L2 Written English","author":"Pradhan A. M.","year":"2010","unstructured":"Pradhan , A. M. , Varde , A. S. , Peng , J. and Fitzpatrick , E. M ., Automatic Classification of Article Errors in L2 Written English . AAAI's FLAIRS , 2010 , pp. 259 -- 264 . Pradhan, A. M., Varde, A. S., Peng, J. and Fitzpatrick, E. M., Automatic Classification of Article Errors in L2 Written English. AAAI's FLAIRS, 2010, pp. 259--264.","journal-title":"AAAI's FLAIRS"},{"key":"e_1_2_1_14_1","first-page":"3209","volume-title":"LREC","author":"Ramos M.A.","year":"2010","unstructured":"Ramos , M.A. , Wanner , L. , Vincze , O. , del Bosque , G. C. , Veiga , N.V. , Su\u00e1rez , E.M. and Gonz\u00e1lez , S.P ., Towards a Motivated Annotation Schema of Collocation Errors in Learner Corpora , LREC , 2010 , pp. 3209 -- 3214 . Ramos, M.A., Wanner, L., Vincze, O., del Bosque, G. C., Veiga, N.V., Su\u00e1rez, E.M. and Gonz\u00e1lez, S.P., Towards a Motivated Annotation Schema of Collocation Errors in Learner Corpora, LREC, 2010, pp. 3209--3214."},{"key":"e_1_2_1_15_1","first-page":"827","article-title":"Corpus-Driven Knowledge Acquisition for Discourse Analysis","author":"Soderland S.","year":"1994","unstructured":"Soderland , S. and Lehnert , W ., Corpus-Driven Knowledge Acquisition for Discourse Analysis , AAAI , 1994 , pp. 827 -- 832 . Soderland, S. and Lehnert, W., Corpus-Driven Knowledge Acquisition for Discourse Analysis, AAAI, 1994, pp. 827--832.","journal-title":"AAAI"},{"key":"e_1_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.websem.2008.06.001"},{"key":"e_1_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.3115\/1073445.1073477"},{"key":"e_1_2_1_18_1","first-page":"155","volume-title":"Demo paper in IEEE International Conference on Information and Communication Systems (ICICS)","author":"Varghese A.","year":"2013","unstructured":"Varghese A. , Varde A. , Peng J. and Fitzpatrick , E. , The CollOrder System for Detecting and Correcting Odd Collocations in L2 Written English , Demo paper in IEEE International Conference on Information and Communication Systems (ICICS) , 2013 , pp. 155 -- 160 . Varghese A., Varde A., Peng J. and Fitzpatrick, E., The CollOrder System for Detecting and Correcting Odd Collocations in L2 Written English, Demo paper in IEEE International Conference on Information and Communication Systems (ICICS), 2013, pp. 155--160."}],"container-title":["ACM SIGKDD Explorations Newsletter"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2830544.2830548","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/2830544.2830548","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T05:43:00Z","timestamp":1750225380000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2830544.2830548"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2015,9,29]]},"references-count":18,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2015,9,29]]}},"alternative-id":["10.1145\/2830544.2830548"],"URL":"https:\/\/doi.org\/10.1145\/2830544.2830548","relation":{},"ISSN":["1931-0145","1931-0153"],"issn-type":[{"type":"print","value":"1931-0145"},{"type":"electronic","value":"1931-0153"}],"subject":[],"published":{"date-parts":[[2015,9,29]]},"assertion":[{"value":"2015-09-29","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}