{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,2,26]],"date-time":"2026-02-26T15:35:07Z","timestamp":1772120107522,"version":"3.50.1"},"reference-count":44,"publisher":"Springer Science and Business Media LLC","issue":"5","license":[{"start":{"date-parts":[[2024,4,23]],"date-time":"2024-04-23T00:00:00Z","timestamp":1713830400000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2024,4,23]],"date-time":"2024-04-23T00:00:00Z","timestamp":1713830400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"name":"ICAR - RENDE"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["SN COMPUT. SCI."],"abstract":"<jats:title>Abstract<\/jats:title>\n                  <jats:p>\n                    In today\u2019s digital era, dominated by social media platforms such as\n                    <jats:italic>Twitter<\/jats:italic>\n                    ,\n                    <jats:italic>Facebook<\/jats:italic>\n                    , and\n                    <jats:italic>Instagram<\/jats:italic>\n                    , the swift dissemination of misinformation represents a significant concern, impacting public sentiment and influencing pivotal global events. Promptly detecting such deceptive content with the help of Machine Learning models is crucial, yet it comes with the challenge of dealing with labelled examples for training these models. Impressive performance results were recently achieved by high-capacity pre-trained transformer-based models (e.g., BERT). Still, such models are too data- and compute-demanding for many critical application contexts where memory, time, and energy consumption must be limited. Here, we propose an innovative semi-supervised method for efficient and effective fake news detection using a content-oriented classifier based on a small-sized BERT embedder. After fine-tuning this model on the sole few labelled data available, an iterative Active Learning (AL) process is carried out, which benefits from limited experts\u2019 feedback to acquire more labelled data for improving the model. The proposed method ensures good detection performances using a few training samples, reasonably small human intervention, and compute\/memory costs.\n                  <\/jats:p>","DOI":"10.1007\/s42979-024-02809-1","type":"journal-article","created":{"date-parts":[[2024,4,23]],"date-time":"2024-04-23T09:02:07Z","timestamp":1713862927000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":12,"title":["Towards Data- and Compute-Efficient Fake-News Detection: An Approach Combining Active Learning and Pre-Trained Language Models"],"prefix":"10.1007","volume":"5","author":[{"given":"Francesco","family":"Folino","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Gianluigi","family":"Folino","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Massimo","family":"Guarascio","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Luigi","family":"Pontieri","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-9119-9865","authenticated-orcid":false,"given":"Paolo","family":"Zicari","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2024,4,23]]},"reference":[{"issue":"5","key":"2809_CR1","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/3395046","volume":"53","author":"X Zhou","year":"2020","unstructured":"Zhou X, Zafarani R. A survey of fake news: fundamental theories, detection methods, and opportunities. ACM Comput Surv. 2020;53(5):1\u201340. https:\/\/doi.org\/10.1145\/3395046.","journal-title":"ACM Comput Surv"},{"key":"2809_CR2","doi-asserted-by":"publisher","first-page":"172","DOI":"10.1007\/978-3-030-29563-9_17","volume-title":"Knowledge science, engineering and management","author":"C Liu","year":"2019","unstructured":"Liu C, Wu X, Yu M, Li G, Jiang J, Huang W, Lu X. A two-stage model based on bert for short fake news detection. In: Douligeris C, Karagiannis D, Apostolou D, editors. Knowledge science, engineering and management. Cham: Springer; 2019. p. 172\u201383."},{"key":"2809_CR3","doi-asserted-by":"publisher","first-page":"133","DOI":"10.1016\/j.aiopen.2022.09.001","volume":"3","author":"L Hu","year":"2022","unstructured":"Hu L, Wei S, Zhao Z, Wu B. Deep learning for fake news detection: a comprehensive survey. AI Open. 2022;3:133\u201355. https:\/\/doi.org\/10.1016\/j.aiopen.2022.09.001.","journal-title":"AI Open"},{"key":"2809_CR4","first-page":"634","volume":"1\u20133","author":"M Guarascio","year":"2018","unstructured":"Guarascio M, Manco G, Ritacco E. Deep learning. Encyclopedia of bioinformatics and computational biology: ABC of bioinformatics. 2018;1\u20133:634\u201347.","journal-title":"Encyclopedia of bioinformatics and computational biology: ABC of bioinformatics"},{"key":"2809_CR5","doi-asserted-by":"crossref","unstructured":"Phan H.T, Nguyen N.T, Hwang D. Fake news detection: a survey of graph neural network methods. Appl Soft Comput. 2023;110235.","DOI":"10.1016\/j.asoc.2023.110235"},{"issue":"2","key":"2809_CR6","doi-asserted-by":"publisher","DOI":"10.1016\/j.ipm.2019.03.004","volume":"57","author":"X Zhang","year":"2020","unstructured":"Zhang X, Ghorbani AA. An overview of online fake news: characterization, detection, and discussion. Informat Process Manage. 2020;57(2): 102025.","journal-title":"Informat Process Manage."},{"key":"2809_CR7","doi-asserted-by":"publisher","DOI":"10.1016\/j.jvcir.2022.103685","volume":"90","author":"C-CJ Kuo","year":"2023","unstructured":"Kuo C-CJ, Madni AM. Green learning: introduction, examples and outlook. J Vis Commun Image Represent. 2023;90: 103685.","journal-title":"J Vis Commun Image Represent"},{"key":"2809_CR8","doi-asserted-by":"publisher","unstructured":"Devlin J, Chang MW, Lee K, Toutanova K. BERT: pre-training of deep bidirectional transformers for language understanding. In: NAACL-HLT, pp. 4171\u20134186, 2019. https:\/\/doi.org\/10.18653\/v1\/N19-1423.","DOI":"10.18653\/v1\/N19-1423"},{"key":"2809_CR9","doi-asserted-by":"crossref","unstructured":"Pelrine K, Danovitch J, Rabbany R. The surprising performance of simple baselines for misinformation detection. In: Proceedings of the Web Conference 2021. WWW \u201921, pp. 3432\u20133441, 2021.","DOI":"10.1145\/3442381.3450111"},{"key":"2809_CR10","doi-asserted-by":"crossref","unstructured":"Guacho GB, Abdali S, Shah N, Papalexakis EE. Semi-supervised content-based detection of misinformation via tensor embeddings. In: Proceedings of the 2018 IEEE\/ACM International Conference on Advances in Social Networks Analysis and Mining. ASONAM \u201918, pp. 322\u2013325, 2020.","DOI":"10.1109\/ASONAM.2018.8508241"},{"key":"2809_CR11","doi-asserted-by":"publisher","unstructured":"Benamira A, Devillers B, Lesot E, Ray AK, Saadi M, Malliaros FD. Semi-supervised learning and graph neural networks for fake news detection. In: ASONAM \u201919: International Conference on Advances in Social Networks Analysis and Mining, Vancouver, British Columbia, Canada, 27-30 August, 2019, pp. 568\u2013569, 2019. https:\/\/doi.org\/10.1145\/3341161.3342958.","DOI":"10.1145\/3341161.3342958"},{"key":"2809_CR12","unstructured":"Meel P, Vishwakarma DK. Fake news detection using semi-supervised graph convolutional network. CoRR abs\/2109.13476 2021; 2109.13476."},{"key":"2809_CR13","doi-asserted-by":"publisher","first-page":"607","DOI":"10.1016\/j.neucom.2021.12.037","volume":"491","author":"SD Das","year":"2022","unstructured":"Das SD, Basak A, Dutta S. A heuristic-driven uncertainty based ensemble framework for fake news detection in tweets and news articles. Neurocomputing. 2022;491:607\u201320. https:\/\/doi.org\/10.1016\/j.neucom.2021.12.037.","journal-title":"Neurocomputing"},{"issue":"14","key":"2809_CR14","doi-asserted-by":"publisher","first-page":"19341","DOI":"10.1007\/s11042-021-11065-x","volume":"81","author":"X Li","year":"2022","unstructured":"Li X, Lu P, Hu L, Wang X, Lu L. A novel self-learning semi-supervised deep learning network to detect fake news on social media. Multimedia Tools Appl. 2022;81(14):19341\u20139. https:\/\/doi.org\/10.1007\/s11042-021-11065-x.","journal-title":"Multimedia Tools Appl"},{"key":"2809_CR15","doi-asserted-by":"publisher","unstructured":"Zicari P, Guarascio M, Pontieri L, Folino G. Learning deep fake-news detectors from scarcely-labelled news corpora. In: Filipe J, Smialek M, Brodsky A, Hammoudi S, editors. Proceedings of the 25th International Conference on Enterprise Information Systems, ICEIS 2023, Vol. 1, Prague, Czech Republic, April 24-26, 2023, pp. 344\u2013353. SCITEPRESS. https:\/\/doi.org\/10.5220\/0011827500003467.","DOI":"10.5220\/0011827500003467"},{"key":"2809_CR16","doi-asserted-by":"crossref","unstructured":"Ren Y, Wang B, Zhang J, Chang Y. Adversarial active learning based heterogeneous graph neural network for fake news detection. In: Proc. of IEEE Intl. Conf. on Data Mining (ICDM\u201920), pp. 452\u2013461, 2020.","DOI":"10.1109\/ICDM50108.2020.00054"},{"key":"2809_CR17","doi-asserted-by":"publisher","DOI":"10.1016\/j.osnem.2023.100244","volume":"33","author":"G Barnab\u00f2","year":"2023","unstructured":"Barnab\u00f2 G, Siciliano F, Castillo C, Leonardi S, Nakov P, Da San Martino G, Silvestri F. Deep active learning for misinformation detection using geometric deep learning. Online Soc Netw Media. 2023;33: 100244.","journal-title":"Online Soc Netw Media"},{"key":"2809_CR18","doi-asserted-by":"crossref","unstructured":"Bhattacharjee SD, Talukder A, Balantrapu BV. Active learning based news veracity detection with feature weighting and deep-shallow fusion. In: Proc. of IEEE Intl. Conf. on Big Data (Big Data\u201917), pp. 556\u2013565, 2017.","DOI":"10.1109\/BigData.2017.8257971"},{"key":"2809_CR19","doi-asserted-by":"crossref","unstructured":"Farinneya P, Pour MMA, Hamidian S, Diab M. Active learning for rumor identification on social media. In: Proc. of Intl. Conf. on Empirical Methods in Natural Language Processing (EMNLP\u201921), pp. 4556\u20134565, 2021.","DOI":"10.18653\/v1\/2021.findings-emnlp.387"},{"key":"2809_CR20","doi-asserted-by":"crossref","unstructured":"Lee K, Mou G, Sievert S. Energy-based domain adaption with active learning for emerging misinformation detection. In: Proc. of IEEE Intl. Conf. on Big Data (Big Data\u201922), pp. 2305\u20132308, 2022.","DOI":"10.1109\/BigData55660.2022.10021038"},{"key":"2809_CR21","volume-title":"Human-in-the-loop machine learning: active learning and annotation for human-centered AI","author":"RM Monarch","year":"2021","unstructured":"Monarch RM. Human-in-the-loop machine learning: active learning and annotation for human-centered AI. USA: Simon and Schuster; 2021."},{"key":"2809_CR22","unstructured":"Sanh V, Debut L, Chaumond J, Wolf T. Distilbert, a distilled version of bert: smaller, faster, cheaper and lighter. arXiv preprint arXiv:1910.01108, 2019."},{"key":"2809_CR23","doi-asserted-by":"crossref","unstructured":"Jiao X, Yin Y, Shang L, Jiang X, Chen X, Li L, Wang F, Liu Q. Tinybert: distilling bert for natural language understanding. In: Findings of the association for computational linguistics: EMNLP 2020, pp. 4163\u20134174.","DOI":"10.18653\/v1\/2020.findings-emnlp.372"},{"key":"2809_CR24","unstructured":"Lan Z, Chen M, Goodman S, Gimpel K, Sharma P, Soricut R. Albert: a lite bert for self-supervised learning of language representations. In: International conference on learning representations. 2020. https:\/\/openreview.net\/forum?id=H1eA7AEtvS."},{"key":"2809_CR25","doi-asserted-by":"crossref","unstructured":"Sun Z, Yu H, Song X, Liu R, Yang Y, Zhou D. Mobilebert: a compact task-agnostic bert for resource-limited devices. In: Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics, pp. 2158\u20132170, 2020.","DOI":"10.18653\/v1\/2020.acl-main.195"},{"key":"2809_CR26","doi-asserted-by":"publisher","first-page":"80207","DOI":"10.1109\/ACCESS.2023.3299858","volume":"11","author":"R Anggrainingsih","year":"2023","unstructured":"Anggrainingsih R, Hassan GM, Datta A. Ce-bert: concise and efficient bert-based model for detecting rumors on twitter. IEEE Access. 2023;11:80207\u201317.","journal-title":"IEEE Access"},{"key":"2809_CR27","unstructured":"Turc I, Chang M-W, Lee K, Toutanova K. Well-read students learn better: On the importance of pre-training compact models. arXiv preprint arXiv:1908.08962, 2019."},{"key":"2809_CR28","unstructured":"Zhou Z, Li L, Chen X, Li A. Mini-Giants: small language models and open source win-win. 2023."},{"key":"2809_CR29","unstructured":"Michel P, Levy O, Neubig G. Are sixteen heads really better than one? Adv Neural Informat Process Syst. 2019;32."},{"key":"2809_CR30","doi-asserted-by":"publisher","DOI":"10.1016\/j.csl.2022.101429","volume":"77","author":"H Sajjad","year":"2023","unstructured":"Sajjad H, Dalvi F, Durrani N, Nakov P. On the effect of dropping layers of pre-trained transformer models. Comput Speech Lang. 2023;77: 101429.","journal-title":"Comput Speech Lang"},{"key":"2809_CR31","doi-asserted-by":"crossref","unstructured":"Anggrainingsih R, Mubashar\u00a0Hassan G, Datta A. Evaluating bert-based pre-training language models for detecting misinformation. arXiv e-prints. 2022;2203.","DOI":"10.21203\/rs.3.rs-1608574\/v1"},{"key":"2809_CR32","unstructured":"Lee D-H. Pseudo-label: the simple and efficient semi-supervised learning method for deep neural networks. ICML 2013 Workshop: challenges in Representation Learning (WREPL). 2013."},{"issue":"9","key":"2809_CR33","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/3472291","volume":"54","author":"P Ren","year":"2021","unstructured":"Ren P, Xiao Y, Chang X, Huang P-Y, Li Z, Gupta BB, Chen X, Wang X. A survey of deep active learning. ACM Comput Surv. 2021;54(9):1\u201340.","journal-title":"ACM Comput Surv"},{"key":"2809_CR34","unstructured":"Dong X, Victor U, Chowdhury S, Qian L. Deep two-path semi-supervised learning for fake news detection. CoRR 2019; 1906.05659 abs\/1906.05659."},{"key":"2809_CR35","doi-asserted-by":"publisher","DOI":"10.1016\/j.eswa.2021.115002","volume":"177","author":"P Meel","year":"2021","unstructured":"Meel P, Vishwakarma DK. A temporal ensembling based semi-supervised convnet for the detection of fake news articles. Exp Syst Appl. 2021;177: 115002. https:\/\/doi.org\/10.1016\/j.eswa.2021.115002.","journal-title":"Exp Syst Appl"},{"key":"2809_CR36","unstructured":"Laine S, Aila T. Temporal ensembling for semi-supervised learning. arXiv preprint arXiv:1610.02242, 2016."},{"key":"2809_CR37","doi-asserted-by":"publisher","unstructured":"Pennington J, Socher R, Manning CD. Glove: global vectors for word representation. In: Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing (EMNLP), pp. 1532\u20131543. https:\/\/doi.org\/10.3115\/v1\/D14-1162 . Association for Computational Linguistics. https:\/\/nlp.stanford.edu\/pubs\/glove.pdf. 2014.","DOI":"10.3115\/v1\/D14-1162"},{"key":"2809_CR38","doi-asserted-by":"crossref","unstructured":"Zicari P, Guarascio M, Pontieri L, Folino G. Learning deep fake-news detectors from scarcely-labelled news corpora. In: Filipe J, Smialek M, Brodsky A, Hammoudi S. editors. Proceedings of the 25th International Conference on Enterprise Information Systems, ICEIS 2023, Vol 1, Prague, Czech Republic, April 24-26,2023, pp. 344\u2013353, 2023.","DOI":"10.5220\/0011827500003467"},{"key":"2809_CR39","doi-asserted-by":"publisher","unstructured":"Devlin J, Chang M-W, Lee K, Toutanova K. Bert: pre-training of deep bidirectional transformers for language understanding. arXiv preprint arXiv:1810.04805 2018. https:\/\/doi.org\/10.18653\/v1\/N19-1423.","DOI":"10.18653\/v1\/N19-1423"},{"issue":"9","key":"2809_CR40","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/3477140","volume":"54","author":"J Mena","year":"2021","unstructured":"Mena J, Pujol O, Vitri\u00e0 J. A survey on uncertainty estimation in deep learning classification systems from a Bayesian perspective. ACM Comput Surv. 2021;54(9):1\u201335.","journal-title":"ACM Comput Surv"},{"issue":"2","key":"2809_CR41","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/3605943","volume":"56","author":"B Min","year":"2023","unstructured":"Min B, Ross H, Sulem E, Veyseh APB, Nguyen TH, Sainz O, Agirre E, Heintz I, Roth D. Recent advances in natural language processing via large pre-trained language models: a survey. ACM Comput Surv. 2023;56(2):1\u201340.","journal-title":"ACM Comput Surv"},{"key":"2809_CR42","unstructured":"Kaplan J, McCandlish S, Henighan T, Brown TB, Chess B, Child R, Gray S, Radford A, Wu J, Amodei D. Scaling laws for neural language models. 2020."},{"key":"2809_CR43","unstructured":"Shu K, Mahudeswaran D, Wang S, Lee D, Liu H. Fakenewsnet: a data repository with news content, social context and dynamic information for studying fake news on social media. arXiv preprint arXiv:1809.01286. 2018."},{"issue":"1","key":"2809_CR44","doi-asserted-by":"publisher","first-page":"22","DOI":"10.1145\/3137597.3137600","volume":"19","author":"K Shu","year":"2017","unstructured":"Shu K, Sliva A, Wang S, Tang J, Liu H. Fake news detection on social media: a data mining perspective. ACM SIGKDD Explor Newsl. 2017;19(1):22\u201336.","journal-title":"ACM SIGKDD Explor Newsl"}],"container-title":["SN Computer Science"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s42979-024-02809-1.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s42979-024-02809-1\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s42979-024-02809-1.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,11,16]],"date-time":"2024-11-16T16:58:24Z","timestamp":1731776304000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s42979-024-02809-1"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,4,23]]},"references-count":44,"journal-issue":{"issue":"5","published-online":{"date-parts":[[2024,6]]}},"alternative-id":["2809"],"URL":"https:\/\/doi.org\/10.1007\/s42979-024-02809-1","relation":{"is-referenced-by":[{"id-type":"doi","id":"10.1007\/s42979-025-04238-0","asserted-by":"object"}]},"ISSN":["2661-8907"],"issn-type":[{"value":"2661-8907","type":"electronic"}],"subject":[],"published":{"date-parts":[[2024,4,23]]},"assertion":[{"value":"30 September 2023","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"19 March 2024","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"23 April 2024","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"Not applicable.","order":2,"name":"Ethics","group":{"name":"EthicsHeading","label":"Ethical approval"}},{"value":"The authors declare that they have no conflict of interests.","order":3,"name":"Ethics","group":{"name":"EthicsHeading","label":"Conflict of interests"}}],"article-number":"470"}}