{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,11,15]],"date-time":"2025-11-15T10:36:11Z","timestamp":1763202971489,"version":"3.40.4"},"publisher-location":"Cham","reference-count":28,"publisher":"Springer Nature Switzerland","isbn-type":[{"type":"print","value":"9783031784972"},{"type":"electronic","value":"9783031784989"}],"license":[{"start":{"date-parts":[[2024,12,4]],"date-time":"2024-12-04T00:00:00Z","timestamp":1733270400000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/www.springernature.com\/gp\/researchers\/text-and-data-mining"},{"start":{"date-parts":[[2024,12,4]],"date-time":"2024-12-04T00:00:00Z","timestamp":1733270400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.springernature.com\/gp\/researchers\/text-and-data-mining"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2025]]},"DOI":"10.1007\/978-3-031-78498-9_2","type":"book-chapter","created":{"date-parts":[[2024,12,3]],"date-time":"2024-12-03T09:26:35Z","timestamp":1733217995000},"page":"16-28","update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":1,"title":[": A New Dataset for\u00a0 pularity diction of\u00a0 manian Reddit Posts"],"prefix":"10.1007","author":[{"given":"Ana-Cristina","family":"Rogoz","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Maria Ilinca","family":"Nechita","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Radu Tudor","family":"Ionescu","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2024,12,4]]},"reference":[{"key":"2_CR1","unstructured":"Almazrouei, E., et\u00a0al.: The falcon series of open language models. arXiv preprint arXiv:2311.16867 (2023)"},{"issue":"1","key":"2_CR2","doi-asserted-by":"publisher","first-page":"21","DOI":"10.1007\/s41109-021-00358-7","volume":"6","author":"K Barnes","year":"2021","unstructured":"Barnes, K., Riesenmy, T., Trinh, M.D., Lleshi, E., Balogh, N., Molontay, R.: Dank or not? Analyzing and predicting the popularity of memes on Reddit. Appl. Netw. Sci. 6(1), 21 (2021)","journal-title":"Appl. Netw. Sci."},{"key":"2_CR3","doi-asserted-by":"publisher","first-page":"135","DOI":"10.1162\/tacl_a_00051","volume":"5","author":"P Bojanowski","year":"2017","unstructured":"Bojanowski, P., Grave, E., Joulin, A., Mikolov, T.: Enriching word vectors with subword information. Trans. Assoc. Comput. Linguist. 5, 135\u2013146 (2017)","journal-title":"Trans. Assoc. Comput. Linguist."},{"issue":"9","key":"2_CR4","doi-asserted-by":"publisher","first-page":"453","DOI":"10.3390\/info11090453","volume":"11","author":"S Carta","year":"2020","unstructured":"Carta, S., Podda, A.S., Recupero, D.R., Saia, R., Usai, G.: Popularity prediction of Instagram posts. Information 11(9), 453 (2020)","journal-title":"Information"},{"key":"2_CR5","doi-asserted-by":"crossref","unstructured":"De, S., Maity, A., Goel, V., Shitole, S., Bhattacharya, A.: Predicting the popularity of Instagram posts for a lifestyle magazine using deep learning. In: Proceedings of CSCITA, pp. 174\u2013177 (2017)","DOI":"10.1109\/CSCITA.2017.8066548"},{"key":"2_CR6","unstructured":"Devlin, J., Chang, M.W., Lee, K., Toutanova, K.: BERT: pre-training of deep bidirectional transformers for language understanding. In: Proceedings of NAACL, pp. 4171\u20134186 (2019)"},{"key":"2_CR7","doi-asserted-by":"crossref","unstructured":"Dumitrescu, S.D., Avram, A.M., Pyysalo, S.: The birth of Romanian BERT. In: Findings of EMNLP, pp. 4324\u20134328 (2020)","DOI":"10.18653\/v1\/2020.findings-emnlp.387"},{"key":"2_CR8","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1016\/j.aiopen.2023.12.002","volume":"5","author":"Z Fang","year":"2024","unstructured":"Fang, Z., et al.: How to generate popular post headlines on social media? AI Open 5, 1\u20139 (2024)","journal-title":"AI Open"},{"key":"2_CR9","doi-asserted-by":"crossref","unstructured":"Ferrer, X., van Nuenen, T., Such, J.M., Criado, N.: Discovering and categorising language biases in Reddit. In: Proceedings of ICWSM, pp. 140\u2013151 (2021)","DOI":"10.1609\/icwsm.v15i1.18048"},{"key":"2_CR10","doi-asserted-by":"crossref","unstructured":"Gjurkovi\u0107, M., \u0160najder, J.: Reddit: a gold mine for personality prediction. In: Proceedings of PEOPLES, pp. 87\u201397 (2018)","DOI":"10.18653\/v1\/W18-1112"},{"key":"2_CR11","doi-asserted-by":"crossref","unstructured":"Hada, R., Sudhir, S., Mishra, P., Yannakoudakis, H., Mohammad, S.M., Shutova, E.: Ruddit: norms of offensiveness for English Reddit comments. In: Proceedings of ACL, pp. 2700\u20132717 (2022)","DOI":"10.18653\/v1\/2021.acl-long.210"},{"key":"2_CR12","doi-asserted-by":"crossref","unstructured":"Joulin, A., Grave, E., Bojanowski, P., Mikolov, T.: Bag of tricks for efficient text classification. arXiv preprint arXiv:1607.01759 (2016)","DOI":"10.18653\/v1\/E17-2068"},{"key":"2_CR13","unstructured":"Kim, J.: Predicting the popularity of reddit posts with AI. arXiv preprint arXiv:2106.07380 (2021)"},{"key":"2_CR14","unstructured":"Kokhlikyan, N., et al.: Captum: a unified and generic model interpretability library for PyTorch. arXiv preprint arXiv:2009.07896 (2020)"},{"key":"2_CR15","unstructured":"Loshchilov, I., Hutter, F.: Decoupled weight decay regularization. In: Proceedings of ICLR (2019)"},{"key":"2_CR16","doi-asserted-by":"publisher","first-page":"1399","DOI":"10.1002\/asi.22844","volume":"64","author":"Z Ma","year":"2013","unstructured":"Ma, Z., Sun, A., Cong, G.: On predicting the popularity of newly emerging hashtags in Twitter. J. Am. Soc. Inform. Sci. Technol. 64, 1399\u20131410 (2013)","journal-title":"J. Am. Soc. Inform. Sci. Technol."},{"key":"2_CR17","doi-asserted-by":"crossref","unstructured":"Mahdavi, M., Asadpour, M., Ghavami, S.: A comprehensive analysis of tweet content and its impact on popularity. In: Proceedings of IST, pp. 559\u2013564 (2016)","DOI":"10.1109\/ISTEL.2016.7881883"},{"key":"2_CR18","doi-asserted-by":"crossref","unstructured":"McHardy, R., Adel, H., Klinger, R.: Adversarial training for satire detection: controlling for confounding variables. In: Proceedings of NAACL, pp. 660\u2013665 (2019)","DOI":"10.18653\/v1\/N19-1069"},{"key":"2_CR19","doi-asserted-by":"crossref","unstructured":"Niculescu, M.A., Ruseti, S., Dascalu, M.: RoGPT2: Romanian GPT2 for text generation. In: Proceedings of ICTAI, pp. 1154\u20131161 (2021)","DOI":"10.1109\/ICTAI52525.2021.00183"},{"key":"2_CR20","doi-asserted-by":"crossref","unstructured":"Poecze, F., Ebster, C., Strauss, C.: Social media metrics and sentiment analysis to evaluate the effectiveness of social media posts. In: Proceedings of ANT-SEIT, pp. 660\u2013666 (2018)","DOI":"10.1016\/j.procs.2018.04.117"},{"issue":"1","key":"2_CR21","first-page":"85","volume":"18","author":"KR Purba","year":"2021","unstructured":"Purba, K.R., Asirvatham, D., Murugesan, R.K.: Instagram post popularity trend analysis and prediction using hashtag, image assessment, and user history features. Int. Arab J. Inf. Technol. 18(1), 85\u201394 (2021)","journal-title":"Int. Arab J. Inf. Technol."},{"key":"2_CR22","doi-asserted-by":"crossref","unstructured":"Shen, J.H., Rudzicz, F.: Detecting anxiety through Reddit. In: Proceedings of CLPsych, pp. 58\u201365 (2017)","DOI":"10.18653\/v1\/W17-3107"},{"key":"2_CR23","doi-asserted-by":"publisher","first-page":"44883","DOI":"10.1109\/ACCESS.2019.2909180","volume":"7","author":"MM Tadesse","year":"2019","unstructured":"Tadesse, M.M., Lin, H., Xu, B., Yang, L.: Detection of depression-related posts in reddit social media forum. IEEE Access 7, 44883\u201344893 (2019)","journal-title":"IEEE Access"},{"key":"2_CR24","doi-asserted-by":"crossref","unstructured":"Turcan, E., McKeown, K.: Dreaddit: a Reddit dataset for stress analysis in social media. In: Proceedings of LOUHI, pp. 97\u2013107 (2019)","DOI":"10.18653\/v1\/D19-6213"},{"issue":"6","key":"2_CR25","doi-asserted-by":"publisher","first-page":"620","DOI":"10.1109\/THMS.2013.2285047","volume":"43","author":"C Wang","year":"2013","unstructured":"Wang, C., Xiao, Z., Liu, Y., Xu, Y., Zhou, A., Zhang, K.: SentiView: sentiment analysis and visualization for Internet popular topics. IEEE Trans. Hum.-Mach. Syst. 43(6), 620\u2013630 (2013)","journal-title":"IEEE Trans. Hum.-Mach. Syst."},{"key":"2_CR26","doi-asserted-by":"crossref","unstructured":"Zhang, Z., Chen, T., Zhou, Z., Li, J., Luo, J.: How to become Instagram famous: post popularity prediction with dual-attention. arXiv preprint arXiv:1809.09314 (2019)","DOI":"10.1109\/BigData.2018.8622461"},{"key":"2_CR27","doi-asserted-by":"crossref","unstructured":"Zhao, Q., Erdogdu, M.A., He, H.Y., Rajaraman, A., Leskovec, J.: SEISMIC: a self-exciting point process model for predicting tweet popularity. In: Proceedings of KDD, pp. 1513\u20131522 (2015)","DOI":"10.1145\/2783258.2783401"},{"key":"2_CR28","unstructured":"Zhu, Y., ul\u00a0Haq, E., Lee, L.H., Tyson, G., Hui, P.: A Reddit dataset for the Russo-Ukrainian conflict in 2022. arXiv preprint arXiv:2206.05107 (2022)"}],"container-title":["Lecture Notes in Computer Science","Pattern Recognition"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/978-3-031-78498-9_2","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,4,27]],"date-time":"2025-04-27T14:15:14Z","timestamp":1745763314000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/978-3-031-78498-9_2"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,12,4]]},"ISBN":["9783031784972","9783031784989"],"references-count":28,"URL":"https:\/\/doi.org\/10.1007\/978-3-031-78498-9_2","relation":{},"ISSN":["0302-9743","1611-3349"],"issn-type":[{"type":"print","value":"0302-9743"},{"type":"electronic","value":"1611-3349"}],"subject":[],"published":{"date-parts":[[2024,12,4]]},"assertion":[{"value":"4 December 2024","order":1,"name":"first_online","label":"First Online","group":{"name":"ChapterHistory","label":"Chapter History"}},{"value":"The data was collected from a publicly available Reddit archive, selecting five Romanian subreddits. The social media posts are freely accessible to the public without any type of subscription. As the data was collected from an archived public website (Reddit), we adhere to the European regulations () that allow researchers to use data in the public web domain for non-commercial research purposes. We thus release our corpus as open-source under a non-commercial share-alike license agreement, namely CC BY-NC-SA 4.0 ().We acknowledge that some posts could refer to certain people, e.g.\u00a0public figures in Romania. Following GDPR regulations, we will remove all references to a person, upon receiving removal requests via an email to any of the authors.","order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Ethics Statement"}},{"value":"ICPR","order":1,"name":"conference_acronym","label":"Conference Acronym","group":{"name":"ConferenceInfo","label":"Conference Information"}},{"value":"International Conference on Pattern Recognition","order":2,"name":"conference_name","label":"Conference Name","group":{"name":"ConferenceInfo","label":"Conference Information"}},{"value":"Kolkata","order":3,"name":"conference_city","label":"Conference City","group":{"name":"ConferenceInfo","label":"Conference Information"}},{"value":"India","order":4,"name":"conference_country","label":"Conference Country","group":{"name":"ConferenceInfo","label":"Conference Information"}},{"value":"2024","order":5,"name":"conference_year","label":"Conference Year","group":{"name":"ConferenceInfo","label":"Conference Information"}},{"value":"1 December 2024","order":7,"name":"conference_start_date","label":"Conference Start Date","group":{"name":"ConferenceInfo","label":"Conference Information"}},{"value":"5 December 2024","order":8,"name":"conference_end_date","label":"Conference End Date","group":{"name":"ConferenceInfo","label":"Conference Information"}},{"value":"27","order":9,"name":"conference_number","label":"Conference Number","group":{"name":"ConferenceInfo","label":"Conference Information"}},{"value":"icpr2024","order":10,"name":"conference_id","label":"Conference ID","group":{"name":"ConferenceInfo","label":"Conference Information"}},{"value":"https:\/\/icpr2024.org\/","order":11,"name":"conference_url","label":"Conference URL","group":{"name":"ConferenceInfo","label":"Conference Information"}}]}}