{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,3,17]],"date-time":"2026-03-17T00:49:38Z","timestamp":1773708578137,"version":"3.50.1"},"reference-count":76,"publisher":"Springer Science and Business Media LLC","issue":"1","license":[{"start":{"date-parts":[[2022,2,10]],"date-time":"2022-02-10T00:00:00Z","timestamp":1644451200000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2022,2,10]],"date-time":"2022-02-10T00:00:00Z","timestamp":1644451200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["J Big Data"],"published-print":{"date-parts":[[2022,12]]},"abstract":"<jats:title>Abstract<\/jats:title><jats:p>Twitter is a frequent target for machine learning research and applications. Many problems, such as sentiment analysis, image tagging, and location prediction have been studied on Twitter data. Much of the prior work that addresses these problems within the context of Twitter focuses on a subset of the types of data available, e.g. only text, or text and image. However, a tweet can have several additional components, such as the location and the author, that can also provide useful information for machine learning tasks. In this work, we explore the problem of jointly modeling several tweet components in a common embedding space via task-agnostic representation learning, which can then be used to tackle various machine learning applications. To address this problem, we propose a deep neural network framework that combines text, image, and graph representations to learn joint embeddings for 5 tweet components: body, hashtags, images, user, and location. In our experiments, we use a large dataset of tweets to learn a joint embedding model and use it in multiple tasks to evaluate its performance vs. state-of-the-art baselines specific to each task. Our results show that our proposed generic method has similar or superior performance to specialized application-specific approaches, including accuracy of 52.43% vs. 48.88% for location prediction and recall of up to 15.93% vs. 12.12% for hashtag recommendation.<\/jats:p>","DOI":"10.1186\/s40537-022-00570-x","type":"journal-article","created":{"date-parts":[[2022,2,10]],"date-time":"2022-02-10T12:02:57Z","timestamp":1644494577000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":9,"title":["Task-agnostic representation learning of multimodal twitter data for downstream applications"],"prefix":"10.1186","volume":"9","author":[{"ORCID":"https:\/\/orcid.org\/0000-0001-5590-0274","authenticated-orcid":false,"given":"Ryan","family":"Rivas","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Sudipta","family":"Paul","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Vagelis","family":"Hristidis","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Evangelos E.","family":"Papalexakis","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Amit K.","family":"Roy-Chowdhury","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2022,2,10]]},"reference":[{"key":"570_CR1","doi-asserted-by":"publisher","first-page":"245","DOI":"10.1016\/j.chb.2019.04.020","volume":"98","author":"D Gruda","year":"2019","unstructured":"Gruda D, Hasan S. Feeling anxious? Perceiving anxiety in tweets using machine learning. Comput Human Behav. 2019;98:245\u201355.","journal-title":"Comput Human Behav"},{"issue":"2","key":"570_CR2","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/2938640","volume":"49","author":"A Giachanou","year":"2016","unstructured":"Giachanou A, Crestani F. Like it or not: a survey of twitter sentiment analysis methods. ACM Comput Surv (CSUR). 2016;49(2):1\u201341.","journal-title":"ACM Comput Surv (CSUR)"},{"key":"570_CR3","doi-asserted-by":"publisher","first-page":"265","DOI":"10.1016\/j.cose.2017.11.013","volume":"76","author":"T Wu","year":"2018","unstructured":"Wu T, Wen S, Xiang Y, Zhou W. Twitter spam detection: survey of new approaches and comparative study. Comput Secur. 2018;76:265\u201384.","journal-title":"Comput Secur"},{"issue":"9","key":"570_CR4","doi-asserted-by":"publisher","first-page":"1652","DOI":"10.1109\/TKDE.2018.2807840","volume":"30","author":"X Zheng","year":"2018","unstructured":"Zheng X, Han J, Sun A. A survey of location prediction on twitter. IEEE Trans Knowl Data Eng. 2018;30(9):1652\u201371.","journal-title":"IEEE Trans Knowl Data Eng"},{"issue":"1","key":"570_CR5","doi-asserted-by":"publisher","first-page":"133","DOI":"10.3390\/s21010133","volume":"21","author":"M Pota","year":"2021","unstructured":"Pota M, Ventura M, Catelli R, Esposito M. An effective bert-based pipeline for twitter sentiment analysis: a case study in italian. Sensors. 2021;21(1):133.","journal-title":"Sensors"},{"key":"570_CR6","doi-asserted-by":"crossref","unstructured":"Chen YC, Lai KT, Liu D, Chen MS. Tagnet: triplet-attention graph networks for hashtag recommendation. IEEE Trans Circuits Syst Video Technol. 2021.","DOI":"10.1109\/TCSVT.2021.3074599"},{"issue":"1","key":"570_CR7","doi-asserted-by":"publisher","first-page":"101121","DOI":"10.1016\/j.joi.2020.101121","volume":"15","author":"MA Masood","year":"2021","unstructured":"Masood MA, Abbasi RA. Using graph embedding and machine learning to identify rebels on twitter. J Informetr. 2021;15(1):101121.","journal-title":"J Informetr"},{"key":"570_CR8","first-page":"103302","volume":"18","author":"Y Liu","year":"2021","unstructured":"Liu Y, Luo X, Zhang M, Tao Z, Liu F. Who are there: discover twitter users and tweets for target area using mention relationship strength and local tweet ratio. J Netw Comput Appl. 2021;18:103302.","journal-title":"J Netw Comput Appl"},{"key":"570_CR9","doi-asserted-by":"crossref","unstructured":"Baek D, Oh Y, Ham B. Exploiting a joint embedding space for generalized zero-shot semantic segmentation. In: Proceedings of ICCV 21, pp. 9536\u20139545; 2021.","DOI":"10.1109\/ICCV48922.2021.00940"},{"key":"570_CR10","doi-asserted-by":"crossref","unstructured":"Qu L, Liu M, Wu J, Gao Z, Nie L. Dynamic modality interaction modeling for image-text retrieval. In: Proceedings of SIGIR \u201821. 2021; pp. 1104\u20131113.","DOI":"10.1145\/3404835.3462829"},{"key":"570_CR11","doi-asserted-by":"crossref","unstructured":"Rawat YS, Kankanhalli MS. Contagnet: Exploiting user context for image tag recommendation. In: Proceedings of MM \u201816. 2016; pp. 1102\u20131106.","DOI":"10.1145\/2964284.2984068"},{"key":"570_CR12","doi-asserted-by":"crossref","unstructured":"Zhang Q, Wang J, Huang H, Huang X, Gong Y. Hashtag recommendation for multimodal microblog using co-attention network. In: Proceedings of IJCAI \u201817. 2017; pp. 3420\u20133426.","DOI":"10.24963\/ijcai.2017\/478"},{"key":"570_CR13","doi-asserted-by":"crossref","unstructured":"Ma R, Qiu X, Zhang Q, Hu X, Jiang YG, Huang X. Co-attention memory network for multimodal microblog\u2019s hashtag recommendation. IEEE Trans Knowl Data Eng. 2019.","DOI":"10.1109\/TKDE.2019.2932406"},{"key":"570_CR14","unstructured":"Faghri F, Fleet DJ, Kiros JR, Fidler S. Vse++: improving visual-semantic embeddings with hard negatives.2017; arXiv:1707.05612."},{"key":"570_CR15","doi-asserted-by":"publisher","first-page":"108153","DOI":"10.1016\/j.patcog.2021.108153","volume":"120","author":"W Zheng","year":"2021","unstructured":"Zheng W, Yin L, Chen X, Ma Z, Liu S, Yang B. Knowledge base graph embedding module design for visual question answering model. Pattern Recogn. 2021;120:108153.","journal-title":"Pattern Recogn"},{"key":"570_CR16","doi-asserted-by":"crossref","unstructured":"Vygon R, Mikhaylovskiy N. Learning efficient representations for keyword spotting with triplet loss.2021; arXiv:2101.04792.","DOI":"10.1007\/978-3-030-87802-3_69"},{"key":"570_CR17","doi-asserted-by":"crossref","unstructured":"Schroff F, Kalenichenko D, Philbin J. Facenet: a unified embedding for face recognition and clustering. In: Proceedings of CVPR \u201815. 2015; pp. 815\u2013823.","DOI":"10.1109\/CVPR.2015.7298682"},{"key":"570_CR18","doi-asserted-by":"crossref","unstructured":"Wu CY, Manmatha R, Smola AJ, Krahenbuhl P. Sampling matters in deep embedding learning. In: Proceedings of ICCV \u201817. 2017; pp. 2840\u20132848.","DOI":"10.1109\/ICCV.2017.309"},{"issue":"2","key":"570_CR19","first-page":"1","volume":"30","author":"Z Zheng","year":"2020","unstructured":"Zheng Z, Zheng L, Garrett M, Yang Y, Xu M, Shen YD. Dual-path convolutional image-text embeddings with instance loss. ACM Trans Multimed Comput Commun Appl (TOMM). 2020;30(2):1\u201323.","journal-title":"ACM Trans Multimed Comput Commun Appl (TOMM)."},{"key":"570_CR20","doi-asserted-by":"crossref","unstructured":"Zhang W, Stratos K. Understanding hard negatives in noise contrastive estimation. 2021; arXiv:2104.06245.","DOI":"10.18653\/v1\/2021.naacl-main.86"},{"key":"570_CR21","doi-asserted-by":"crossref","unstructured":"Mithun NC, Panda R, Papalexakis EE, Roy-Chowdhury AK. Webly supervised joint embedding for cross-modal image-text retrieval. In: Proceedings of MM \u201818. 2018; pp. 1856\u20131864.","DOI":"10.1145\/3240508.3240712"},{"key":"570_CR22","doi-asserted-by":"crossref","unstructured":"Wang Z, Liu X, Li H, Sheng L, Yan J, Wang X, Shao J. Camp: cross-modal adaptive message passing for text-image retrieval. In: Proceedings of the 2019 IEEE International Conference on Computer Vision. 2019; pp. 5764\u20135773.","DOI":"10.1109\/ICCV.2019.00586"},{"key":"570_CR23","doi-asserted-by":"crossref","unstructured":"Lee KH, Chen X, Hua G, Hu H, He X. Stacked cross attention for image-text matching. In: Proceedings of ECCV \u201818. 2018; pp. 201\u2013216.","DOI":"10.1007\/978-3-030-01225-0_13"},{"key":"570_CR24","doi-asserted-by":"crossref","unstructured":"Mithun NC, Li J, Metze F, Roy-Chowdhury AK. Learning joint embedding with multimodal cues for cross-modal video-text retrieval. In: Proceedings of ICMR \u201818. 2018; pp. 19\u201327.","DOI":"10.1145\/3206025.3206064"},{"key":"570_CR25","doi-asserted-by":"crossref","unstructured":"Dong J, Li X, Xu C, Ji S, He Y, Yang G, Wang X. Dual encoding for zero-example video retrieval. In: Proceedings of CVPR \u201819. 2019; pp. 9346\u20139355.","DOI":"10.1109\/CVPR.2019.00957"},{"key":"570_CR26","doi-asserted-by":"crossref","unstructured":"Wray M, Larlus D, Csurka G, Damen D. Fine-grained action retrieval through multiple parts-of-speech embeddings. In: Proceedings of ICCV \u201819. 2019; pp. 450\u2013459.","DOI":"10.1109\/ICCV.2019.00054"},{"key":"570_CR27","unstructured":"Liu Y, Albanie S, Nagrani A, Zisserman A. Use what you have: video retrieval using representations from collaborative experts. 2019; arXiv:1907.13487."},{"key":"570_CR28","doi-asserted-by":"crossref","unstructured":"Yu Y, Kim J, Kim G. A joint sequence fusion model for video question answering and retrieval. In: Proceedings of ECCV \u201818. 2018; pp. 471\u2013487.","DOI":"10.1007\/978-3-030-01234-2_29"},{"key":"570_CR29","doi-asserted-by":"crossref","unstructured":"Zhang B, Hu H, Sha F. Cross-modal and hierarchical modeling of video and text. In: Proceedings of ECCV \u201818.2018; pp. 374\u2013390.","DOI":"10.1007\/978-3-030-01261-8_23"},{"key":"570_CR30","doi-asserted-by":"crossref","unstructured":"Shao D, Xiong Y, Zhao Y, Huang Q, Qiao Y, Lin D. Find and focus: retrieve and localize video events with natural language queries. In: Proceedings of ECCV \u201818. 2018; pp. 200\u2013216.","DOI":"10.1007\/978-3-030-01240-3_13"},{"key":"570_CR31","doi-asserted-by":"crossref","unstructured":"Hendricks LA, Wang O, Shechtman E, Sivic J, Darrell T, Russell B. Localizing moments in video with natural language. In: Proceedings of ICCV \u201817. 2017; pp. 5803\u20135812.","DOI":"10.1109\/ICCV.2017.618"},{"key":"570_CR32","unstructured":"Escorcia V, Soldan M, Sivic J, Ghanem B, Russell B. Temporal localization of moments in video collections with natural language. 2019; arXiv:1907.12763."},{"key":"570_CR33","doi-asserted-by":"crossref","unstructured":"Paul S, Mithun NC, Roy-Chowdhury AK. Text-based localization of moments in a video corpus. 2020; arXiv:2008.08716.","DOI":"10.1109\/TIP.2021.3120038"},{"key":"570_CR34","unstructured":"Hahn M, Silva A, Rehg JM. Action2vec: a crossmodal embedding approach to action learning. 2019;arXiv:1901.00484."},{"key":"570_CR35","doi-asserted-by":"crossref","unstructured":"Zhu D, Ma Y, Liu Y. Deepad: a joint embedding approach for anomaly detection on attributed networks. In: Proceedings of ICCS \u201820. 2020; pp. 294\u2013307.","DOI":"10.1007\/978-3-030-50417-5_22"},{"key":"570_CR36","doi-asserted-by":"crossref","unstructured":"Li C, Cao Y, Hou L, Shi J, Li J, Chua TS. Semi-supervised entity alignment via joint knowledge embedding model and cross-graph model. In: Proceedings of EMNLP-IJCNLP \u201819 .2019; pp. 2723\u20132732.","DOI":"10.18653\/v1\/D19-1274"},{"issue":"6","key":"570_CR37","doi-asserted-by":"publisher","first-page":"1987","DOI":"10.1007\/s00521-020-05088-y","volume":"33","author":"B Xiong","year":"2020","unstructured":"Xiong B, Bao P, Wu Y. Learning semantic and relationship joint embedding for author name disambiguation. Neural Comput Appl. 2020;33(6):1987\u201398.","journal-title":"Neural Comput Appl"},{"key":"570_CR38","doi-asserted-by":"crossref","unstructured":"Dhingra B, Zhou Z, Fitzpatrick D, Muehl M, Cohen WW. Tweet2vec: character-based distributed representations for social media. 2016; arXiv:1605.03481.","DOI":"10.18653\/v1\/P16-2044"},{"key":"570_CR39","doi-asserted-by":"crossref","unstructured":"Vosoughi S, Vijayaraghavan P, Roy D. Tweet2vec: learning tweet embeddings using character-level cnn-lstm encoder-decoder. In: Proceedings of SIGIR \u201816.2016; pp. 1041\u20131044.","DOI":"10.1145\/2911451.2914762"},{"key":"570_CR40","unstructured":"M\u00fcller M, Salath\u00e9 M, Kummervold PE. Covid-twitter-bert: a natural language processing model to analyse COVID-19 content on twitter. 2020; arXiv:2005.07503."},{"key":"570_CR41","first-page":"1097","volume":"25","author":"A Krizhevsky","year":"2012","unstructured":"Krizhevsky A, Sutskever I, Hinton GE. Imagenet classification with deep convolutional neural networks. Adv Neural Inform Process Syst. 2012;25:1097\u2013105.","journal-title":"Adv Neural Inform Process Syst"},{"key":"570_CR42","unstructured":"Simonyan K, Zisserman A. Very deep convolutional networks for large-scale image recognition. 2014; arXiv:1409.1556."},{"issue":"1","key":"570_CR43","doi-asserted-by":"publisher","first-page":"102435","DOI":"10.1016\/j.ipm.2020.102435","volume":"58","author":"RK Behera","year":"2021","unstructured":"Behera RK, Jena M, Rath SK, Misra S. Co-lstm: convolutional lstm model for sentiment analysis in social big data. Inform Process Manage. 2021;58(1):102435.","journal-title":"Inform Process Manage."},{"key":"570_CR44","unstructured":"Lu J, Yang J, Batra D, Parikh D. Hierarchical question-image co-attention for visual question answering. In: Proceedings of NIPS \u201816. 2016; pp. 289\u2013297."},{"key":"570_CR45","doi-asserted-by":"crossref","unstructured":"Davis CA, Varol O, Ferrara E, Flammini A, Menczer F. Botornot: a system to evaluate social bots. In: Proceedings of WWW \u201816 Companion. 2016; pp. 273\u2013274.","DOI":"10.1145\/2872518.2889302"},{"issue":"1","key":"570_CR46","doi-asserted-by":"publisher","first-page":"5","DOI":"10.1023\/A:1010933404324","volume":"45","author":"L Breiman","year":"2001","unstructured":"Breiman L. Random forests. Mach Learn. 2001;45(1):5\u201332.","journal-title":"Mach Learn."},{"key":"570_CR47","doi-asserted-by":"crossref","unstructured":"Sayyadiharikandeh M, Varol O, Yang KC, Flammini A, Menczer F. Detection of novel social bots by ensembles of specialized classifiers. In: Proceedings of CIKM \u201820. 2020; pp. 2725\u20132732.","DOI":"10.1145\/3340531.3412698"},{"issue":"1","key":"570_CR48","doi-asserted-by":"publisher","first-page":"48","DOI":"10.1002\/hbe2.115","volume":"1","author":"KC Yang","year":"2019","unstructured":"Yang KC, Varol O, Davis CA, Ferrara E, Flammini A, Menczer F. Arming the public with artificial intelligence to counter social bots. Human Behav Emerg Technol. 2019;1(1):48\u201361.","journal-title":"Human Behav Emerg Technol"},{"key":"570_CR49","doi-asserted-by":"publisher","first-page":"312","DOI":"10.1016\/j.ins.2018.08.019","volume":"467","author":"S Kudugunta","year":"2018","unstructured":"Kudugunta S, Ferrara E. Deep neural networks for bot detection. Inform Sci. 2018;467:312\u201322.","journal-title":"Inform Sci."},{"key":"570_CR50","doi-asserted-by":"crossref","unstructured":"Matsuo S, Shimoda W, Yanai K. Twitter photo geo-localization using both textual and visual features. In: Proceedings of BigMM \u201817. 2017; pp. 22\u201325.","DOI":"10.1109\/BigMM.2017.74"},{"issue":"11","key":"570_CR51","doi-asserted-by":"publisher","first-page":"4503","DOI":"10.1007\/s12652-018-1134-3","volume":"10","author":"S Kumar","year":"2019","unstructured":"Kumar S, Nezhurina MI. An ensemble classification approach for prediction of user\u2019s next location based on twitter data. J Amb Intel Human Comput. 2019;10(11):4503\u201313.","journal-title":"J Amb Intel Human Comput."},{"key":"570_CR52","unstructured":"Lau JH, Chi L, Tran KN, Cohn T. End-to-end network for twitter geolocation prediction and hashing. In: Proceedings of IJCNLP \u201817. 2017; pp. 744\u2013753."},{"key":"570_CR53","doi-asserted-by":"publisher","first-page":"121","DOI":"10.1186\/s40537-021-00503-0","volume":"8","author":"K Linnell","year":"2021","unstructured":"Linnell K, Arnold M, Alshaabi T, McAndrew T, Lim J, Dodds PS, Danforth CM. The sleep loss insult of spring daylight savings in the us is observable in twitter activity. J Big Data. 2021;8:121.","journal-title":"J Big Data."},{"key":"570_CR54","doi-asserted-by":"publisher","first-page":"72","DOI":"10.1186\/s40537-021-00463-5","volume":"8","author":"A Feizollah","year":"2021","unstructured":"Feizollah A, Mostafa MM, Sulaiman A, Zakaria Z, Firdaus A. Exploring halal tourism tweets on social media. J Big Data. 2021;8:72.","journal-title":"J Big Data."},{"key":"570_CR55","doi-asserted-by":"publisher","first-page":"65","DOI":"10.1186\/s40537-019-0228-x","volume":"6","author":"CA Pi\u00f1a-Garc\u00eda","year":"2019","unstructured":"Pi\u00f1a-Garc\u00eda CA, Ram\u00edrez-Ram\u00edrez L. Exploring crime patterns in Mexico city. J Big Data. 2019;6:65.","journal-title":"J Big Data."},{"key":"570_CR56","doi-asserted-by":"crossref","unstructured":"He K, Zhang X, Ren S, Sun J. Deep residual learning for image recognition. In: Proceedings of CVPR \u201816. 2016; pp. 770\u2013778.","DOI":"10.1109\/CVPR.2016.90"},{"key":"570_CR57","doi-asserted-by":"crossref","unstructured":"Lin T, Maire M, Belongie S, Hays J, Perona P, Ramanan D, Doll\u00e1r P, Zitnick CL. Microsoft coco: common objects in context. In: Proceedings of ECCV \u201814. 2014; pp. 740\u2013755.","DOI":"10.1007\/978-3-319-10602-1_48"},{"key":"570_CR58","doi-asserted-by":"publisher","first-page":"67","DOI":"10.1162\/tacl_a_00166","volume":"2","author":"P Young","year":"2014","unstructured":"Young P, Lai A, Hodosh M, Hockenmaier J. From image descriptions to visual denotations: new similarity metrics for semantic inference over event descriptions. Trans Assoc Comput Linguist. 2014;2:67\u201378.","journal-title":"Trans Assoc Comput Linguist."},{"key":"570_CR59","first-page":"8026","volume":"32","author":"A Paszke","year":"2019","unstructured":"...Paszke A, Gross S, Massa F, Lerer A, Bradbury J, Chanan G, Killeen T, Lin Z, Gimelshein N, Antiga L, Desmaison A, K\u00f6pf A, Yang E, DeVito Z, Raison M, Tejani A, Chilamkurthy S, Steiner B, Fang L, Bai J, Chintala S. Pytorch: an imperative style, high-performance deep learning library. Adv Neural Inform Process Syst. 2019;32:8026\u201337.","journal-title":"Adv Neural Inform Process Syst."},{"key":"570_CR60","doi-asserted-by":"publisher","first-page":"135","DOI":"10.1162\/tacl_a_00051","volume":"5","author":"P Bojanowski","year":"2017","unstructured":"Bojanowski P, Grave E, Joulin A, Mikolov T. Enriching word vectors with subword information. Trans Assoc Comput Linguist. 2017;5:135\u201346.","journal-title":"Trans Assoc Comput Linguist."},{"key":"570_CR61","doi-asserted-by":"crossref","unstructured":"Deng J, Dong W, Socher R, Li LJ, Li K, Fei-Fei L. Imagenet: a large-scale hierarchical image database. In: Proceedings of CVPR \u201809. 2009; pp. 248\u2013255.","DOI":"10.1109\/CVPR.2009.5206848"},{"key":"570_CR62","unstructured":"Abraham L. fastnode2vec (2020). https:\/\/doi.org\/10.5281\/zenodo.3902632 Accessed Accessed 26 Mar 2021."},{"key":"570_CR63","doi-asserted-by":"crossref","unstructured":"Grover A, Leskovec J. Node2vec: scalable feature learning for networks. In: Proceedings of SIGKDD \u201816. 2016; pp. 855\u2013864.","DOI":"10.1145\/2939672.2939754"},{"key":"570_CR64","unstructured":"Kingma DP, Ba J. Adam: a method for stochastic optimization. 2014; arXiv: 1412.6980."},{"issue":"2","key":"570_CR65","doi-asserted-by":"publisher","first-page":"114","DOI":"10.1016\/j.jides.2016.10.001","volume":"3","author":"S Giannoulakis","year":"2016","unstructured":"Giannoulakis S, Tsapatsoulis N. Evaluating the descriptive power of instagram hashtags. J Innov Digital Ecosyst. 2016;3(2):114\u201329.","journal-title":"J Innov Digital Ecosyst."},{"key":"570_CR66","unstructured":"Lee K, Eoff B, Caverlee J. Seven months with the devils: a long-term study of content polluters on twitter. In: Proceedings of ICWSM \u201811 (2011)."},{"key":"570_CR67","doi-asserted-by":"crossref","unstructured":"Varol O, Ferrara E, Davis C, Menczer F. Flammini, A. Online human\u2013bot interactions: detection, estimation, and characterization. In: Proceedings of ICWSM \u201817.2017.","DOI":"10.1609\/icwsm.v11i1.14871"},{"key":"570_CR68","doi-asserted-by":"crossref","unstructured":"Cresci, S., Di\u00a0Pietro, R., Petrocchi, M., Spognardi, A., Tesconi, M. The paradigm-shift of social spambots: evidence, theories, and tools for the arms race. In: Proceedings of WWW \u201817 Companion; 2017. pp. 963\u2013972.","DOI":"10.1145\/3041021.3055135"},{"key":"570_CR69","doi-asserted-by":"crossref","unstructured":"Gilani Z, Farahbakhsh R, Tyson G, Wang L, Crowcroft J. Of bots and humans (on twitter). In: Proceedings of ASONAM \u201817. 2017; pp. 349\u2013354.","DOI":"10.1145\/3110025.3110090"},{"key":"570_CR70","doi-asserted-by":"crossref","unstructured":"Mazza M, Cresci S, Avvenuti M, Quattrociocchi W, Tesconi M. Rtbust: exploiting temporal patterns for botnet detection on twitter. In: Proceedings of WEBSCI \u201819, 2019; pp. 183\u2013192.","DOI":"10.1145\/3292522.3326015"},{"key":"570_CR71","doi-asserted-by":"crossref","unstructured":"Cresci S, Lillo F, Regoli D, Tardelli S, Tesconi M. \\$ fake: evidence of spam and bot activity in stock microblogs on twitter. In: Proceedings of ICWSM \u201818. 2018.","DOI":"10.1609\/icwsm.v12i1.15073"},{"key":"570_CR72","doi-asserted-by":"crossref","unstructured":"Yang KC, Varol O, Hui PM, Menczer F. Scalable and generalizable social bot detection through data selection. In: Proceedings of AAAI \u201820, 2020; pp. 1096\u20131103.","DOI":"10.1609\/aaai.v34i01.5460"},{"issue":"10","key":"570_CR73","doi-asserted-by":"publisher","first-page":"0241045","DOI":"10.1371\/journal.pone.0241045","volume":"15","author":"A Rauchfleisch","year":"2020","unstructured":"Rauchfleisch A, Kaiser J. The false positive problem of automatic bot detection in social science research. PloS One. 2020;15(10):0241045.","journal-title":"PloS One."},{"key":"570_CR74","doi-asserted-by":"crossref","unstructured":"Kim Y. Convolutional neural networks for sentence classification. In: Proceedings of EMNLP \u201814. 2014; pp. 1746\u20131751.","DOI":"10.3115\/v1\/D14-1181"},{"key":"570_CR75","doi-asserted-by":"publisher","first-page":"556","DOI":"10.35940\/ijeat.E9666.069520","volume":"9","author":"RH Singh","year":"2020","unstructured":"Singh RH, Maurya S, Tripathi T, Narula T, Srivastav G. Movie recommendation system using cosine similarity and knn. Int J Eng Adv Technol. 2020;9:556\u20139.","journal-title":"Int J Eng Adv Technol."},{"key":"570_CR76","first-page":"2825","volume":"12","author":"F Pedregosa","year":"2011","unstructured":"Pedregosa F, Varoquaux G, Gramfort A, Michel V, Thirion B, Grisel O, Blondel M, Prettenhofer P, Weiss R, Dubourg V, Vanderplas J, Passos A, Cournapeau D, Brucher M, Perrot M, Duchesnay E. Scikit-learn: machine learning in python. J Mach Learn Res. 2011;12:2825\u201330.","journal-title":"J Mach Learn Res."}],"container-title":["Journal of Big Data"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1186\/s40537-022-00570-x.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1186\/s40537-022-00570-x\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1186\/s40537-022-00570-x.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,9,18]],"date-time":"2024-09-18T07:19:52Z","timestamp":1726643992000},"score":1,"resource":{"primary":{"URL":"https:\/\/journalofbigdata.springeropen.com\/articles\/10.1186\/s40537-022-00570-x"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,2,10]]},"references-count":76,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2022,12]]}},"alternative-id":["570"],"URL":"https:\/\/doi.org\/10.1186\/s40537-022-00570-x","relation":{},"ISSN":["2196-1115"],"issn-type":[{"value":"2196-1115","type":"electronic"}],"subject":[],"published":{"date-parts":[[2022,2,10]]},"assertion":[{"value":"2 October 2021","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"26 January 2022","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"10 February 2022","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"Not applicable.","order":2,"name":"Ethics","group":{"name":"EthicsHeading","label":"Ethics approval and consent to participate"}},{"value":"Not applicable.","order":3,"name":"Ethics","group":{"name":"EthicsHeading","label":"Consent for publication"}},{"value":"The authors declare that they have no competing interests.","order":4,"name":"Ethics","group":{"name":"EthicsHeading","label":"Competing interests"}}],"article-number":"18"}}