{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,14]],"date-time":"2026-05-14T07:15:00Z","timestamp":1778742900891,"version":"3.51.4"},"reference-count":24,"publisher":"Association for Computing Machinery (ACM)","issue":"8","license":[{"start":{"date-parts":[[2023,8,23]],"date-time":"2023-08-23T00:00:00Z","timestamp":1692748800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Asian Low-Resour. Lang. Inf. Process."],"published-print":{"date-parts":[[2023,8,31]]},"abstract":"<jats:p>Idiomatic expressions are important natural parts of all languages and prominent parts of our daily speech. Idioms cannot be interpreted from the words that they are formed with directly and people may not understand the meaning. From past literature, it was noted that idiom affects Natural Language Processing research like machine translation, semantic analysis, and sentiment analysis. Other languages like English, Chinese, and Indian idioms are recognized through different methods in different research. As there is no standard method and research to identify Amharic idioms, this study is aimed to build a model to identify idioms for the Amharic language using a supervised machine learning approach. The study used 800 labeled expressions for training and 200 expressions for testing from Amharic idiom books \u201c\u12e8\u12a0\u121b \u1228\u129b \u1348\u120a\u1326\u127d\u201d and different Amharic documents. To measure the performance of the model, we used accuracy, precision, recall, and F-score. Finally, a 97.5% accuracy result was achieved from the testing dataset showing a promising result. The study contributes to the information systems discourse about improving the awareness and knowledge of researchers on Amharic idioms.<\/jats:p>","DOI":"10.1145\/3606864","type":"journal-article","created":{"date-parts":[[2023,7,3]],"date-time":"2023-07-03T12:09:11Z","timestamp":1688386151000},"page":"1-9","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":5,"title":["Automatic Idiom Identification Model for Amharic Language"],"prefix":"10.1145","volume":"22","author":[{"ORCID":"https:\/\/orcid.org\/0009-0001-1853-6984","authenticated-orcid":false,"given":"Anduamlak","family":"Abebe Fenta","sequence":"first","affiliation":[{"name":"Debre Tabor University, Computer Science Department, Ethiopia"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-9974-5015","authenticated-orcid":false,"given":"Seffi","family":"Gebeyehu","sequence":"additional","affiliation":[{"name":"Bahirdar Institute of Technology, Bahirdar University, Ethiopia"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2023,8,23]]},"reference":[{"key":"e_1_3_1_2_2","volume-title":"\u12e8\u12a0\u121b\u1228\u129b\u1348\u120a\u1326\u127d (Amharic Idioms","author":"Aklilu A.","year":"1992","unstructured":"A. Aklilu and D. Worku. 1992. \u12e8\u12a0\u121b\u1228\u129b\u1348\u120a\u1326\u127d (Amharic Idioms) (2nd edition ed.). Mega Publishing Agency, Addis Ababa.","edition":"2"},{"key":"e_1_3_1_3_2","first-page":"395","article-title":"Monolingual alignment by edit rate computation on sentential paraphrase pairs","volume":"2","author":"Bouamor Houda","year":"2011","unstructured":"Houda Bouamor, Aur\u00e9lien Max, and Anne Vilnat. 2011. Monolingual alignment by edit rate computation on sentential paraphrase pairs. ACL-HLT 2011 - Proc. 49th Annu. Meet. Assoc. Comput. Linguist. Hum. Lang. Technol. 2, (2011), 395\u2013400.","journal-title":"ACL-HLT 2011 - Proc. 49th Annu. Meet. Assoc. Comput. Linguist. Hum. Lang. Technol"},{"key":"e_1_3_1_4_2","doi-asserted-by":"publisher","DOI":"10.3115\/1698239.1698243"},{"key":"e_1_3_1_5_2","doi-asserted-by":"publisher","DOI":"10.1515\/jisys-2020-0021"},{"key":"e_1_3_1_6_2","doi-asserted-by":"publisher","DOI":"10.1162\/coli.08-010-R1-07-048"},{"key":"e_1_3_1_7_2","first-page":"2908","volume-title":"Proc. 7th Int. Conf. Lang. Resour. Eval. Lr. 2010","author":"Fritzinger Fabienne","year":"2010","unstructured":"Fabienne Fritzinger, Marion Weller, and Ulrich Heid. 2010. A survey of idiomatic preposition-noun-verb triples on token level. Proc. 7th Int. Conf. Lang. Resour. Eval. Lr. 2010 (2010), 2908\u20132914."},{"key":"e_1_3_1_8_2","article-title":"Vector representations of text data in deep learning","author":"Grzegorczyk Karol","year":"2019","unstructured":"Karol Grzegorczyk. 2019. Vector representations of text data in deep learning. AGH University of Science and Technology. Retrieved from http:\/\/arxiv.org\/abs\/1901.01695.","journal-title":"AGH University of Science and Technology"},{"key":"e_1_3_1_9_2","doi-asserted-by":"publisher","DOI":"10.3115\/1273073.1273119"},{"issue":"2","key":"e_1_3_1_10_2","first-page":"5","article-title":"Information extraction system for Amharic text","volume":"5","author":"Hirpassa Sinatyehu","year":"2017","unstructured":"Sinatyehu Hirpassa. 2017. Information extraction system for Amharic text. Int. J. Comput. Sci. Trends Technol. 5, 2 (2017), 5\u201315.","journal-title":"Int. J. Comput. Sci. Trends Technol."},{"key":"e_1_3_1_11_2","doi-asserted-by":"publisher","DOI":"10.22046\/LA.2017.22"},{"key":"e_1_3_1_12_2","first-page":"1417","volume-title":"EMNLP 2013 - 2013 Conf. Empir. Methods Nat. Lang. Process. Proc. Conf. October","author":"Muzny Grace","year":"2013","unstructured":"Grace Muzny and Luke Zettlemoyer. 2013. Automatic idiom identification in Wiktionary. EMNLP 2013 - 2013 Conf. Empir. Methods Nat. Lang. Process. Proc. Conf. October (2013), 1417\u20131421."},{"key":"e_1_3_1_13_2","doi-asserted-by":"publisher","DOI":"10.14445\/22312803\/ijctt-v48p126"},{"key":"e_1_3_1_14_2","first-page":"17","volume-title":"Research in Computing Science, Special issue: Natural Language Processing and its Applications","author":"Peng J.","year":"2010","unstructured":"J. Peng, Anna Feldman, and L. Street. 2010. Computing linear discriminants for idiomatic sentence detection. In Research in Computing Science, Special issue: Natural Language Processing and its Applications, The Center for Computing Research of IPN, Mexico, 17\u201328."},{"key":"e_1_3_1_15_2","volume-title":"2nd Annual International Symposium, SIMBig 2015","author":"Peng Jing","year":"2015","unstructured":"Jing Peng and Anna Feldman. 2015. Automatic idiom recognition with word embeddings. In 2nd Annual International Symposium, SIMBig 2015. Springer, USA, New Jersey."},{"key":"e_1_3_1_16_2","first-page":"2752","volume-title":"COLING 2016 - 26th Int. Conf. Comput. Linguist. Proc. COLING 2016 Tech. Pap","author":"Peng Jing","year":"2016","unstructured":"Jing Peng and Anna Feldman. 2016. Experiments in idiom recognition. COLING 2016 - 26th Int. Conf. Comput. Linguist. Proc. COLING 2016 Tech. Pap. (2016), 2752\u20132761."},{"key":"e_1_3_1_17_2","doi-asserted-by":"publisher","DOI":"10.21427\/D77H8K"},{"key":"e_1_3_1_18_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/p16-1019"},{"key":"e_1_3_1_19_2","doi-asserted-by":"publisher","DOI":"10.26615\/978-954-452-049-6-083"},{"key":"e_1_3_1_20_2","doi-asserted-by":"publisher","DOI":"10.1109\/ISCC.2006.27"},{"key":"e_1_3_1_21_2","first-page":"681","volume-title":"Int. Conf. Recent Adv. Nat. Lang. Process. RANLP 2015-January","author":"Verma Rakesh","year":"2015","unstructured":"Rakesh Verma and Vasanthi Vuppuluri. 2015. A new approach for idiom identification using meanings and the web. Int. Conf. Recent Adv. Nat. Lang. Process. RANLP 2015-January (2015), 681\u2013687."},{"key":"e_1_3_1_22_2","first-page":"697","volume-title":"Proc. 5th Int. Conf. Lang. Resour. Eval. Lr","author":"Vilar David","year":"2006","unstructured":"David Vilar, Jia Xu, Luis Fernando D'Haro, and Hermann Ney. 2006. Error analysis of statistical machine translation output. Proc. 5th Int. Conf. Lang. Resour. Eval. Lr. (2006), 697\u2013702."},{"key":"e_1_3_1_23_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.eswa.2015.05.039"},{"key":"e_1_3_1_24_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2020.coling-main.91"},{"key":"e_1_3_1_25_2","volume-title":"Proceedings of the First International Conference on Web Research (ICWR), Tehran, Iran (15\u201316)","author":"Zolfaghari Mayam","year":"2013","unstructured":"Mayam Zolfaghari, Vahiden Nazeri, Fatemeh Sefidkon, and Farhad Rejali. 2013. A survey on Twitter sentiment analysis. In Proceedings of the First International Conference on Web Research (ICWR), Tehran, Iran (15\u201316)."}],"container-title":["ACM Transactions on Asian and Low-Resource Language Information Processing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3606864","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3606864","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T17:48:52Z","timestamp":1750182532000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3606864"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,8,23]]},"references-count":24,"journal-issue":{"issue":"8","published-print":{"date-parts":[[2023,8,31]]}},"alternative-id":["10.1145\/3606864"],"URL":"https:\/\/doi.org\/10.1145\/3606864","relation":{},"ISSN":["2375-4699","2375-4702"],"issn-type":[{"value":"2375-4699","type":"print"},{"value":"2375-4702","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,8,23]]},"assertion":[{"value":"2022-03-31","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2023-06-18","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2023-08-23","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}