{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,9]],"date-time":"2026-07-09T04:23:40Z","timestamp":1783571020492,"version":"3.55.0"},"reference-count":35,"publisher":"MDPI AG","issue":"4","license":[{"start":{"date-parts":[[2023,12,14]],"date-time":"2023-12-14T00:00:00Z","timestamp":1702512000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"National Natural Science Foundation of China","award":["62276216"],"award-info":[{"award-number":["62276216"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["BDCC"],"abstract":"<jats:p>Spain possesses a vast number of poems. Most have features that mean they present significantly different styles. A superficial reading of these poems may confuse readers due to their complexity. Therefore, it is of vital importance to classify the style of the poems in advance. Currently, poetry classification studies are mostly carried out manually, which creates extremely high requirements for the professional quality of classifiers and consumes a large amount of time. Furthermore, the objectivity of the classification cannot be guaranteed because of the influence of the classifier\u2019s subjectivity. To solve these problems, a Spanish poetry classification framework was designed using artificial intelligence technology, which improves the accuracy, efficiency, and objectivity of classification. First, an artificial-intelligence-driven Spanish poetry classification framework is described in detail, and is illustrated by a framework diagram to clearly represent each step in the process. The framework includes many algorithms and models, such as the Term Frequency\u2013Inverse Document Frequency (TF_IDF), Bagging, Support Vector Machines (SVMs), Adaptive Boosting (AdaBoost), logistic regression (LR), Gradient Boosting Decision Trees (GBDT), LightGBM (LGB), eXtreme Gradient Boosting (XGBoost), and Random Forest (RF). The roles of each algorithm in the framework are clearly defined. Finally, experiments were performed for model selection, comparing the results of these algorithms.The Bagging model stood out for its high accuracy, and the experimental results showed that the proposed framework can help researchers carry out poetry research work more efficiently, accurately, and objectively.<\/jats:p>","DOI":"10.3390\/bdcc7040183","type":"journal-article","created":{"date-parts":[[2023,12,15]],"date-time":"2023-12-15T03:16:33Z","timestamp":1702610193000},"page":"183","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":8,"title":["An Artificial-Intelligence-Driven Spanish Poetry Classification Framework"],"prefix":"10.3390","volume":"7","author":[{"given":"Shutian","family":"Deng","sequence":"first","affiliation":[{"name":"School of Hispanic and Portuguese Studies, Beijing Foreign Studies University, Beijing 100089, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Gang","family":"Wang","sequence":"additional","affiliation":[{"name":"School of Computing and Artificial Intelligence, Southwest Jiaotong University, Chengdu 610031, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Hongjun","family":"Wang","sequence":"additional","affiliation":[{"name":"School of Computing and Artificial Intelligence, Southwest Jiaotong University, Chengdu 610031, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Fuliang","family":"Chang","sequence":"additional","affiliation":[{"name":"School of Hispanic and Portuguese Studies, Beijing Foreign Studies University, Beijing 100089, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"1968","published-online":{"date-parts":[[2023,12,14]]},"reference":[{"key":"ref_1","unstructured":"Cavnar, W.B., and Trenkle, J.M. (1994, January 11\u201313). N-gram-based text categorization. Proceedings of the SDAIR-94, 3rd Annual Symposium on Document Analysis and Information Retrieval, Las Vegas, NV, USA."},{"key":"ref_2","doi-asserted-by":"crossref","unstructured":"Lewis, D.D. (1992, January 23\u201326). Feature selection and feature extraction for text categorization. Proceedings of the Speech and Natural Language: Proceedings of a Workshop Held at Harriman, New York, NY, USA.","DOI":"10.3115\/1075527.1075574"},{"key":"ref_3","doi-asserted-by":"crossref","first-page":"61","DOI":"10.14257\/ijdta.2014.7.1.06","article-title":"KNN based machine learning approach for text and document mining","volume":"7","author":"Bijalwan","year":"2014","journal-title":"Int. J. Database Theory Appl."},{"key":"ref_4","doi-asserted-by":"crossref","unstructured":"Larkey, L.S., and Croft, W.B. (1996, January 18\u201322). Combining classifiers in text categorization. Proceedings of the 19th Annual International ACM SIGIR Conference on Research and Development in Information Retrieval, Zurich, Switzerland.","DOI":"10.1145\/243199.243276"},{"key":"ref_5","doi-asserted-by":"crossref","first-page":"843","DOI":"10.1126\/science.267.5199.843","article-title":"Gauging similarity with n-grams: Language-independent categorization of text","volume":"267","author":"Damashek","year":"1995","journal-title":"Science"},{"key":"ref_6","doi-asserted-by":"crossref","first-page":"400","DOI":"10.1007\/s10791-008-9083-7","article-title":"Using the Web as corpus for self-training text categorization","volume":"12","author":"Rosso","year":"2009","journal-title":"Inf. Retr."},{"key":"ref_7","doi-asserted-by":"crossref","first-page":"110","DOI":"10.1016\/j.knosys.2018.03.003","article-title":"An Automated Text Categorization Framework based on Hyperparameter Optimization","volume":"149","author":"Tellez","year":"2017","journal-title":"Knowl.-Based Syst."},{"key":"ref_8","unstructured":"Barbado, A., Gonz\u00e1lez, M.D., and Carrera, D. (2021). Lexico-semantic and affective modelling of Spanish poetry: A semi-supervised learning approach. arXiv."},{"key":"ref_9","doi-asserted-by":"crossref","first-page":"258","DOI":"10.1002\/asi.24532","article-title":"A bridge too far for artificial intelligence?: Automatic classification of stanzas in Spanish poetry","volume":"73","author":"Rosa","year":"2022","journal-title":"J. Assoc. Inf. Sci. Technol."},{"key":"ref_10","doi-asserted-by":"crossref","first-page":"15","DOI":"10.3389\/fdigh.2018.00015","article-title":"On Poetic Topic Modeling: Extracting Themes and Motifs From a Corpus of Spanish Poetry","volume":"5","author":"Borja","year":"2018","journal-title":"Front. Digit. Humanit."},{"key":"ref_11","doi-asserted-by":"crossref","first-page":"438","DOI":"10.3390\/info12110438","article-title":"Emotion Classification in Spanish: Exploring the Hard Classes","volume":"12","author":"Chiruzzo","year":"2021","journal-title":"Information"},{"key":"ref_12","doi-asserted-by":"crossref","unstructured":"Barros, L., Rodriguez, P., and Ortigosa, A. (2013, January 2\u20135). Automatic Classification of Literature Pieces by Emotion Detection: A Study on Quevedo\u2019s Poetry. Proceedings of the 2013 Humaine Association Conference on Affective Computing and Intelligent Interaction, IEEE, Geneva, Switzerland.","DOI":"10.1109\/ACII.2013.30"},{"key":"ref_13","doi-asserted-by":"crossref","first-page":"112","DOI":"10.1093\/llc\/fqx009","article-title":"A metrical scansion system for fixed-metre Spanish poetry","volume":"33","year":"2018","journal-title":"Digit. Scholarsh. Humanit."},{"key":"ref_14","doi-asserted-by":"crossref","unstructured":"Torres-Moreno, J.M., and Moreno-Jim\u00e9nez, L.G. (2020). LiSSS: A toy corpus of Spanish Literary Sentences for Emotions detection. arXiv.","DOI":"10.13053\/cys-24-3-3474"},{"key":"ref_15","first-page":"2723","article-title":"Marathi poem classification using machine learning","volume":"8","author":"Deshmukh","year":"2019","journal-title":"Int. J. Recent Technol. Eng."},{"key":"ref_16","unstructured":"Ara\u00fajo, P., and Mamede, N. (2023, October 14). Classificador de Poemas. In Proceedings of the Confer\u00eancia Cient\u00edfica e Tecnol\u00f3gica em Engenharia. Available online: https:\/\/www.hlt.inesc-id.pt\/documents\/papers\/2002Araujo.pdf."},{"key":"ref_17","doi-asserted-by":"crossref","first-page":"1701","DOI":"10.11591\/eei.v9i4.1898","article-title":"English poems categorization using text mining and rough set theory","volume":"9","author":"Alsaidi","year":"2020","journal-title":"Bull. Electr. Eng. Inform."},{"key":"ref_18","doi-asserted-by":"crossref","first-page":"40","DOI":"10.1524\/glot.2013.0014","article-title":"Automatic categorization of ottoman poems","volume":"4","author":"Can","year":"2013","journal-title":"Glottotheory"},{"key":"ref_19","doi-asserted-by":"crossref","unstructured":"Zhu, M., Wang, G., Li, C., Wang, H., and Zhang, B. (2023). Artificial Intelligence Classification Model for Modern Chinese Poetry in Education. Sustainability, 15.","DOI":"10.3390\/su15065265"},{"key":"ref_20","doi-asserted-by":"crossref","unstructured":"Kaur, J., and Saini, J.R. (2017, January 29\u201331). Punjabi poetry classification: The test of 10 machine learning algorithms. Proceedings of the 9th International Conference on Machine Learning and Computing, Hong Kong, China.","DOI":"10.1145\/3055635.3056589"},{"key":"ref_21","first-page":"358","article-title":"Gujarati poetry classification based on emotions using deep learning","volume":"6","author":"Mehta","year":"2021","journal-title":"Int. J. Eng. Appl. Sci. Technol."},{"key":"ref_22","doi-asserted-by":"crossref","unstructured":"de la Rosa, J., P\u00e9rez, \u00c1., Hern, L., Ros, S., and Gonz, E. (2020, January 5\u20137). PoetryLab as Infrastructure for the Analysis of Spanish Poetry. Proceedings of the CLARIN Annual Conference, Virtual.","DOI":"10.3384\/ecp1809"},{"key":"ref_23","doi-asserted-by":"crossref","first-page":"51734","DOI":"10.1109\/ACCESS.2021.3069635","article-title":"Automated metric analysis of Spanish poetry: Two complementary approaches","volume":"9","author":"Marco","year":"2021","journal-title":"IEEE Access"},{"key":"ref_24","doi-asserted-by":"crossref","unstructured":"Zhao, K., Huang, L., Song, R., Shen, Q., and Xu, H. (2021). A sequential graph neural network for short text classification. Algorithms, 14.","DOI":"10.3390\/a14120352"},{"key":"ref_25","doi-asserted-by":"crossref","unstructured":"Huang, Y., Song, R., Giunchiglia, F., and Xu, H. (2022). A multitask learning framework for abuse detection and emotion classification. Algorithms, 15.","DOI":"10.3390\/a15040116"},{"key":"ref_26","doi-asserted-by":"crossref","unstructured":"Papadia, G., Pacella, M., and Giliberti, V. (2022). Topic Modeling for Automatic Analysis of Natural Language: A Case Study in an Italian Customer Support Center. Algorithms, 15.","DOI":"10.3390\/a15060204"},{"key":"ref_27","doi-asserted-by":"crossref","unstructured":"Campos Macias, N., D\u00fcggelin, W., Ruf, Y., and Hanne, T. (2022). Building a technology recommender system using web crawling and natural language processing Technology. Algorithms, 15.","DOI":"10.3390\/a15080272"},{"key":"ref_28","doi-asserted-by":"crossref","unstructured":"Neagu, D.C., Rus, A.B., Grec, M., Boroianu, M.A., Bogdan, N., and Gal, A. (2022). Towards Sentiment Analysis for Romanian Twitter Content. Algorithms, 15.","DOI":"10.3390\/a15100357"},{"key":"ref_29","doi-asserted-by":"crossref","unstructured":"Tang, H., Kamei, S., and Morimoto, Y. (2023). Data Augmentation Methods for Enhancing Robustness in Text Classification Tasks. Algorithms, 16.","DOI":"10.3390\/a16010059"},{"key":"ref_30","doi-asserted-by":"crossref","unstructured":"Zhang, X., Zhou, H., Yu, K., Wu, X., and Yazidi, A. (2023). Tsetlin Machine for Sentiment Analysis and Spam Review Detection in Chinese. Algorithms, 16.","DOI":"10.3390\/a16020093"},{"key":"ref_31","doi-asserted-by":"crossref","unstructured":"Liu, H., Ye, Z., Zhao, H., and Yang, Y. (2023). Chinese Text De-Colloquialization Technique Based on Back-Translation Strategy and End-to-End Learning. Appl. Sci., 13.","DOI":"10.3390\/app131910818"},{"key":"ref_32","doi-asserted-by":"crossref","unstructured":"Torres-Silva, E.A., R\u00faa, S., Giraldo-Forero, A.F., Durango, M.C., Fl\u00f3rez-Arango, J.F., and Orozco-Duque, A. (2023). Classification of Severe Maternal Morbidity from Electronic Health Records Written in Spanish Using Natural Language Processing. Appl. Sci., 13.","DOI":"10.3390\/app131910725"},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"Li, J., and Wu, C. (2023). Deep Learning and Text Mining: Classifying and Extracting Key Information from Construction Accident Narratives. Appl. Sci., 13.","DOI":"10.3390\/app131910599"},{"key":"ref_34","doi-asserted-by":"crossref","unstructured":"Ahn, S. (2023). Experimental Study of Morphological Analyzers for Topic Categorization in News Articles. Appl. Sci., 13.","DOI":"10.3390\/app131910572"},{"key":"ref_35","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/3458754","article-title":"Domain-specific language model pretraining for biomedical natural language processing","volume":"3","author":"Gu","year":"2021","journal-title":"ACM Trans. Comput. Healthc."}],"container-title":["Big Data and Cognitive Computing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/2504-2289\/7\/4\/183\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,10]],"date-time":"2025-10-10T21:39:06Z","timestamp":1760132346000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/2504-2289\/7\/4\/183"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,12,14]]},"references-count":35,"journal-issue":{"issue":"4","published-online":{"date-parts":[[2023,12]]}},"alternative-id":["bdcc7040183"],"URL":"https:\/\/doi.org\/10.3390\/bdcc7040183","relation":{},"ISSN":["2504-2289"],"issn-type":[{"value":"2504-2289","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,12,14]]}}}