{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,10]],"date-time":"2026-07-10T13:27:34Z","timestamp":1783690054132,"version":"3.55.0"},"reference-count":35,"publisher":"MDPI AG","issue":"1","license":[{"start":{"date-parts":[[2021,1,7]],"date-time":"2021-01-07T00:00:00Z","timestamp":1609977600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Information"],"abstract":"<jats:p>The need to fight the progressive negative impact of fake news is escalating, which is evident in the strive to do research and develop tools that could do this job. However, a lack of adequate datasets and good word embeddings have posed challenges to make detection methods sufficiently accurate. These resources are even totally missing for \u201clow-resource\u201d African languages, such as Amharic. Alleviating these critical problems should not be left for tomorrow. Deep learning methods and word embeddings contributed a lot in devising automatic fake news detection mechanisms. Several contributions are presented, including an Amharic fake news detection model, a general-purpose Amharic corpus (GPAC), a novel Amharic fake news detection dataset (ETH_FAKE), and Amharic fasttext word embedding (AMFTWE). Our Amharic fake news detection model, evaluated with the ETH_FAKE dataset and using the AMFTWE, performed very well.<\/jats:p>","DOI":"10.3390\/info12010020","type":"journal-article","created":{"date-parts":[[2021,1,10]],"date-time":"2021-01-10T19:55:56Z","timestamp":1610308556000},"page":"20","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":42,"title":["Combating Fake News in \u201cLow-Resource\u201d Languages: Amharic Fake News Detection Accompanied by Resource Crafting"],"prefix":"10.3390","volume":"12","author":[{"given":"Fantahun","family":"Gereme","sequence":"first","affiliation":[{"name":"Institute of Fundamental and Frontier Sciences, University of Electronic Science and Technology of China, Chengdu 611731, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"William","family":"Zhu","sequence":"additional","affiliation":[{"name":"Institute of Fundamental and Frontier Sciences, University of Electronic Science and Technology of China, Chengdu 611731, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-9413-6459","authenticated-orcid":false,"given":"Tewodros","family":"Ayall","sequence":"additional","affiliation":[{"name":"School of Computer Science and Engineering, University of Electronic Science and Technology of China, Chengdu 611731, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Dagmawi","family":"Alemu","sequence":"additional","affiliation":[{"name":"School of Computer Science and Engineering, University of Electronic Science and Technology of China, Chengdu 611731, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"1968","published-online":{"date-parts":[[2021,1,7]]},"reference":[{"key":"ref_1","unstructured":"Bajaj, S. (2017). The Pope Has a New Baby! Fake News Detection Using Deep Learning, Stanford University."},{"key":"ref_2","doi-asserted-by":"crossref","unstructured":"Ferreira, W., and Vlachos, A. (2016, January 12\u201317). Emergent: A novel data-set for stance classification. Proceedings of the 2016 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, San Diego, CA, USA.","DOI":"10.18653\/v1\/N16-1138"},{"key":"ref_3","unstructured":"Ma, J., Gao, W., Mitra, P., Kwon, S., Jansen, B.J., Wong, K., and Cha, M. (2016, January 9\u201315). Detecting rumors from microblogs with recurrent neural networks. Proceedings of the Twenty-Fifth International Joint Conference on Artificial Intelligence (IJCAI-16), New York, NY, USA."},{"key":"ref_4","doi-asserted-by":"crossref","first-page":"22","DOI":"10.1145\/3137597.3137600","article-title":"Fake news detection on social media: A data mining perspective","volume":"19","author":"Shu","year":"2017","journal-title":"SIGKDD"},{"key":"ref_5","doi-asserted-by":"crossref","unstructured":"Zhou, X., and Zafarani, R. (2020). 2018 A survey of fake news: Fundamental theories, detection methods, and opportunities. ACM Comput. Surv., 53.","DOI":"10.1145\/3395046"},{"key":"ref_6","unstructured":"Gereme, F.B., and William, Z. (2020, January 13\u201315). Fighting fake news using deep learning: Pre-trained word embeddings and the embedding layer investigated. Proceedings of the 3rd International Conference on Computational Intelligence and Intelligent Systems (CIIS 2020), Tokyo, Japan."},{"key":"ref_7","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/3003433","article-title":"Stance and sentiment in tweets","volume":"17","author":"Mohammad","year":"2017","journal-title":"ACM Trans. Internet Technol."},{"key":"ref_8","unstructured":"(2021, January 06). District of Columbia Language Access Act Fact Sheet 2004, Available online: https:\/\/ohr.dc.gov\/publication\/know-your-rights-cards-amharic."},{"key":"ref_9","doi-asserted-by":"crossref","unstructured":"Karimi, H., and Tanh, J. (2019, January 2\u20137). Learning hierarchical discourse-level structure for fake news detection. Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics, Minneapolis, MN, USA.","DOI":"10.18653\/v1\/N19-1347"},{"key":"ref_10","unstructured":"Khan, J.Y., Khondaker, T.I., Iqbal, A., and Afroz, S. (2019). A benchmark study on machine learning methods for fake news detection. arXiv."},{"key":"ref_11","doi-asserted-by":"crossref","unstructured":"Singhania, S., Fernandez, N., and Rao, S. (2017, January 14\u201318). 3HAN: A deep neural network for fake news detection. Proceedings of the International Conference on Neural Information Processing, Guangzhou, China.","DOI":"10.1007\/978-3-319-70096-0_59"},{"key":"ref_12","unstructured":"Tagami, T., Ouchi, H., Asano, H., Hanawa, K., Uchiyama, K., Suzuki, K., Inui, K., Komiya, A., Fujimura, A., and Yanai, H. (2018, January 1\u20133). Suspicious news detection using micro blog text. Proceedings of the 32nd Pacific Asia Conference on Language, Information and Computation, Hong Kong, China."},{"key":"ref_13","first-page":"10","article-title":"Fake news detection: A deep learning approach","volume":"1","author":"Thota","year":"2018","journal-title":"SMU Data Sci. Rev."},{"key":"ref_14","doi-asserted-by":"crossref","unstructured":"Wang, W.Y. (2017, January 6). Liar-Liar pants on fire: A new benchmark dataset for fake news detection. Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers), Vancouver, BC, Canada.","DOI":"10.18653\/v1\/P17-2067"},{"key":"ref_15","unstructured":"Yang, Y., Zheng, L., Zhang, J., Cui, Q., Li, Z., and Yu, P. (2018). TI-CNN: Convolutional neural networks for fake news detection. arXiv."},{"key":"ref_16","first-page":"117","article-title":"Amharic as lingua franca in Ethiopia","volume":"20","author":"Ronny","year":"2006","journal-title":"Lissan J. Afr. Lang. Linguist."},{"key":"ref_17","doi-asserted-by":"crossref","unstructured":"Rosenhouse, J., and Kowner, R. (2008). Amharic: Political and social effects on English loan words. Globally Speaking: Motives for Adopting English Vocabulary in Other Languages, Multilingual Matters.","DOI":"10.21832\/9781847690524"},{"key":"ref_18","first-page":"1","article-title":"Manual annotation of Amharic news items with part-of-speech tags and its challenges","volume":"Volume 2","author":"Demeke","year":"2016","journal-title":"Ethiopian Languages Research Center Working Papers"},{"key":"ref_19","doi-asserted-by":"crossref","unstructured":"Rychl\u00fd, P., and Suchomel, V. (2016, January 12\u201316). Annotated Amharic Corpora. Proceedings of the International Conference on Text, Speech, and Dialogue(TSD2016), Brno, Czech Republic.","DOI":"10.1007\/978-3-319-45510-5_34"},{"key":"ref_20","first-page":"1","article-title":"The Cr\u00fabad\u00e1n Project: Corpus building for under-resourced languages","volume":"5","author":"Scannell","year":"2007","journal-title":"Cah. Cental"},{"key":"ref_21","unstructured":"Gezmu, A.M., Seyoum, B.E., Gasser, M., and N\u00fcrnberger, A. (2018). Contemporary Amharic Corpus: Automatically Morpho-Syntactically Tagged Amharic Corpus. Proceedings of the First Workshop on Linguistic Resources for Natural Language Processing, Santa Fe, NM, USA, 2018, Association for Computational Linguistics."},{"key":"ref_22","doi-asserted-by":"crossref","unstructured":"Gereme, F.B., and William, Z. (2019, January 28\u201330). Early detection of fake news, before it flies high. Proceedings of the 2nd International Conference on Big Data Technologies (ICBDT2019), Jinan, China.","DOI":"10.1145\/3358528.3358567"},{"key":"ref_23","doi-asserted-by":"crossref","unstructured":"Tang, D., Wei, F., Yang, N., Zhou, M., Liu, T., and Qin, B. (2014, January 23\u201325). Learning sentiment-specific word embedding for twitter sentiment classification. Proceedings of the 52nd Annual Meeting of the Association for Computational Linguistics 1, Baltimore, MD, USA.","DOI":"10.3115\/v1\/P14-1146"},{"key":"ref_24","doi-asserted-by":"crossref","first-page":"2059","DOI":"10.1109\/TASLP.2016.2598310","article-title":"Transition-Based dependency parsing exploiting supertags","volume":"24","author":"Ouchi","year":"2016","journal-title":"IEEE\/ACM Trans. Audio Speech Lang. Process."},{"key":"ref_25","doi-asserted-by":"crossref","first-page":"266","DOI":"10.1109\/TASLP.2017.2772846","article-title":"A neural approach to source dependence based context model for statistical machine translation","volume":"26","author":"Chen","year":"2018","journal-title":"IEEE\/ACM Trans. Audio Speech Lang. Process."},{"key":"ref_26","unstructured":"Grave, E., Bojanowski, P., Gupta, P., Joulin, A., and Mikolov, T. (2018, January 7\u201312). Learning word vectors for 157 languages. Proceedings of the Eleventh International Conference on Language Resources and Evaluation (LREC 2018), Miyazaki, Japan."},{"key":"ref_27","unstructured":"Mikolov, T., Chen, K., Corrado, G., and Dean, J. (2013, January 2\u20134). Efficient estimation of word representations in vector space. Proceedings of the International Conference on Learning Representations (ICLR 2013), Scottsdale, AZ, USA."},{"key":"ref_28","unstructured":"Mikolov, T., Sutskever, I., Chen, K., Corrado, G., and Dean, J. (2013, January 5\u201310). Distributed representations of words and phrases and their compositionality. Proceedings of the 26th International Conference on Neural Information Processing Systems\u2014Volume 2, Lake Tahoe, SN, USA."},{"key":"ref_29","doi-asserted-by":"crossref","unstructured":"Pennington, J., Socher, R., and Manning, C.D. (2014, January 25\u201329). GloVe-Global Vectors for word representation. Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing (EMNLP), Doha, Qatar.","DOI":"10.3115\/v1\/D14-1162"},{"key":"ref_30","doi-asserted-by":"crossref","first-page":"135","DOI":"10.1162\/tacl_a_00051","article-title":"Enriching word vectors with subword information","volume":"5","author":"Bojanowski","year":"2017","journal-title":"TACL"},{"key":"ref_31","unstructured":"Bairong, Z., Wenbo, W., Zhiyu, L., Chonghui, Z., and Shinozaki, T. (2017, January 10). Comparative analysis of word embedding methods for DSTC6 end-to-end Conversation Modeling Track. Proceedings of the 6th Dialog System Technology Challenges (DSTC6) Workshop, Long Beach, CA, USA."},{"key":"ref_32","unstructured":"Li, H., Li, X., Caragea, D., and Caragea, C. (2018, January 4\u20137). Comparison of word embeddings and sentence encodings as generalized representations for crisis tweet classification tasks. Proceedings of the ISCRAM Asian Pacific Conference, Wellington, New Zealand."},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"Wang, B., Wang, A., Chen, F., Wang, Y., and Kuo, C.J. (2019). Evaluating word embedding models: Methods and experimental results. APSIPA Trans. Signal Inf. Process., 8.","DOI":"10.1017\/ATSIP.2019.12"},{"key":"ref_34","doi-asserted-by":"crossref","unstructured":"Schnabel, T., Labutov, I., Mimno, D., and Joachims, T. (2015, January 17\u201321). Evaluation methods for unsupervised word embeddings. Proceedings of the 2015 Conference on Empirical Methods in Natural Language Processing, Lisbon, Portugal.","DOI":"10.18653\/v1\/D15-1036"},{"key":"ref_35","doi-asserted-by":"crossref","first-page":"211","DOI":"10.1257\/jep.31.2.211","article-title":"Social media and fake news in the 2016 election","volume":"31","author":"Allcott","year":"2017","journal-title":"J. Econ. Perspect."}],"container-title":["Information"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/2078-2489\/12\/1\/20\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,11]],"date-time":"2025-10-11T05:08:09Z","timestamp":1760159289000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/2078-2489\/12\/1\/20"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2021,1,7]]},"references-count":35,"journal-issue":{"issue":"1","published-online":{"date-parts":[[2021,1]]}},"alternative-id":["info12010020"],"URL":"https:\/\/doi.org\/10.3390\/info12010020","relation":{},"ISSN":["2078-2489"],"issn-type":[{"value":"2078-2489","type":"electronic"}],"subject":[],"published":{"date-parts":[[2021,1,7]]}}}