{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,3,27]],"date-time":"2025-03-27T12:17:19Z","timestamp":1743077839570,"version":"3.40.3"},"publisher-location":"Singapore","reference-count":30,"publisher":"Springer Nature Singapore","isbn-type":[{"type":"print","value":"9789811692284"},{"type":"electronic","value":"9789811692291"}],"license":[{"start":{"date-parts":[[2022,1,1]],"date-time":"2022-01-01T00:00:00Z","timestamp":1640995200000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2022,1,1]],"date-time":"2022-01-01T00:00:00Z","timestamp":1640995200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2022]]},"abstract":"<jats:title>Abstract<\/jats:title><jats:p>With the increasing of the telecom network fraud in China, SMS (Short Message Service) has became an important channel exploited by the criminals to contact victims. Due to the tiny amount compared with normal SMS, the high proportion of malicious adversarial characters, and the lack of knowledge to specific fraud types, it is still challenging to identify the fraud SMS efficiently. In this paper, we firstly conduct a measurement study to explore the characteristics of the fraud SMS. Based on the exploration, we propose a two-stage algorithm called TFC. TFC can quickly filter out normal SMS in the first stage with two indicator functions, and then easily identifies the category of fraud SMS in the second stage by combining the semantic deep features and the domain-knowledge based artificial features. We conduct two real-world SMS datasets for extensive experiments, and the results show that TFC successfully reduces calculation cost and achieves better performance in distinguishing various categories of fraud SMS.<\/jats:p>","DOI":"10.1007\/978-981-16-9229-1_10","type":"book-chapter","created":{"date-parts":[[2022,1,21]],"date-time":"2022-01-21T12:03:56Z","timestamp":1642766636000},"page":"157-175","update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":1,"title":["TFC: Defending Against SMS Fraud via a Two-Stage Algorithm"],"prefix":"10.1007","author":[{"given":"Gaoxiang","family":"Li","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yuzhong","family":"Ye","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yangfei","family":"Shi","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jinlin","family":"Chen","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Dexing","family":"Chen","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Anyang","family":"Li","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2022,1,21]]},"reference":[{"key":"10_CR1","unstructured":"Supreme People's Procuratorate sets up a steering group to crackdown the rapidly spread cybercrime. http:\/\/www.xinhuanet.com\/legal\/2020-04\/08\/c_1125829699.htm. Accessed 07 Apr 2021"},{"key":"10_CR2","unstructured":"The survey report shows that three telecom operators have achieved nearly 100% scam SMS processing in September 2020. https:\/\/www.sohu.com\/a\/430647655_287480. Accessed 07 Apr 2021"},{"key":"10_CR3","unstructured":"China public security department cracked 256,000 cases of telecommunications network fraud in 2020. http:\/\/www.gov.cn\/xinwen\/2021-01\/02\/content_5576223.htm. Accessed 07 Apr 2021"},{"key":"10_CR4","unstructured":"Wikipedia, PINYIN. https:\/\/en.wikipedia.org\/wiki\/Pinyin. Accessed 07 Apr 2021"},{"key":"10_CR5","doi-asserted-by":"crossref","unstructured":"Cui, Y., et al.: Revisiting pre-trained models for Chinese natural language processing. In: Proceedings of Conference on Empirical Methods in Natural Language Processing: Findings (2020)","DOI":"10.18653\/v1\/2020.findings-emnlp.58"},{"key":"10_CR6","unstructured":"DATASET Repo. https:\/\/github.com\/leegdcert\/TFCDATASET. Accessed 07 Apr 2021"},{"issue":"1","key":"10_CR7","doi-asserted-by":"publisher","first-page":"29","DOI":"10.1007\/s11235-016-0269-9","volume":"66","author":"JW Joo","year":"2017","unstructured":"Joo, J.W., Moon, S.Y., Singh, S., Park, J.H.: S-Detector: an enhanced security model for detecting smishing attack for mobile computing. Telecommun. Syst. 66(1), 29\u201338 (2017). https:\/\/doi.org\/10.1007\/s11235-016-0269-9","journal-title":"Telecommun. Syst."},{"key":"10_CR8","doi-asserted-by":"publisher","first-page":"502","DOI":"10.1007\/978-981-10-8660-1_38","volume-title":"Smart and Innovative Trends in Next Generation Computing Technologies","author":"D Goel","year":"2018","unstructured":"Goel, D., Jain, A.: Smishing-classifier: a novel framework for detection of smishing attack in mobile environment. In: Bhattacharyya, P., Sastry, H.G., Marriboyina, V., Sharma, R. (eds.) Smart and Innovative Trends in Next Generation Computing Technologies, pp. 502\u2013512. Springer, Singapore (2018). https:\/\/doi.org\/10.1007\/978-981-10-8660-1_38"},{"key":"10_CR9","doi-asserted-by":"publisher","first-page":"803","DOI":"10.1016\/j.future.2020.03.021","volume":"108","author":"S Mishra","year":"2020","unstructured":"Mishra, S., Soni, D.: Smishing detector: a security model to detect smishing through SMS content analysis and URL behavior analysis. Futur. Gener. Comput. Syst. 108, 803\u2013815 (2020)","journal-title":"Futur. Gener. Comput. Syst."},{"key":"10_CR10","doi-asserted-by":"crossref","unstructured":"Pervaiz, F., et al.: An assessment of SMS fraud in Pakistan. In: Proceedings of 2nd ACM SIGCAS Conference on Computing and Sustainable Societies (2019)","DOI":"10.1145\/3314344.3332500"},{"issue":"10","key":"10_CR11","doi-asserted-by":"publisher","first-page":"9899","DOI":"10.1016\/j.eswa.2012.02.053","volume":"39","author":"S Delany","year":"2012","unstructured":"Delany, S., Buckley, M., Greene, D.: SMS spam filtering: methods and data. Expert Syst. Appl. 39(10), 9899\u20139908 (2012)","journal-title":"Expert Syst. Appl."},{"key":"10_CR12","doi-asserted-by":"publisher","first-page":"15650","DOI":"10.1109\/ACCESS.2017.2666785","volume":"5","author":"S Abdulhamid","year":"2017","unstructured":"Abdulhamid, S., et al.: A review on mobile SMS spam filtering techniques. IEEE Access 5, 15650\u201315666 (2017)","journal-title":"IEEE Access"},{"key":"10_CR13","doi-asserted-by":"publisher","first-page":"246","DOI":"10.1016\/j.knosys.2011.08.018","volume":"26","author":"D Olszewski","year":"2012","unstructured":"Olszewski, D.: A probabolistic approch to fraud detection in telecommunications. Knowl.-Based Syst. 26, 246\u2013258 (2012)","journal-title":"Knowl.-Based Syst."},{"issue":"1","key":"10_CR14","doi-asserted-by":"publisher","first-page":"182","DOI":"10.1016\/j.engappai.2010.05.009","volume":"24","author":"H Farvaresh","year":"2011","unstructured":"Farvaresh, H., Sepehri, M.: A data mining framework for detecting subscription fraud in telecommunication. Eng. Appl. Artif. Intell. 24(1), 182\u2013194 (2011)","journal-title":"Eng. Appl. Artif. Intell."},{"key":"10_CR15","doi-asserted-by":"crossref","unstructured":"Wang, Y., Bansal, M.: Robust machine comprehension models via adversarial training. In: Proceedings of NAACL-HLT (2018)","DOI":"10.18653\/v1\/N18-2091"},{"key":"10_CR16","unstructured":"Ebrahimi, J., Lowd, D., Dou, D.: On adversarial examples for character-level neural machine translation. In: Proceedings of 27th International Conference on Computational Linguistics (2018)"},{"key":"10_CR17","doi-asserted-by":"crossref","unstructured":"Li, J., Ji, S., Du, T., Li, B., Wang, T.: TEXTBUGGER: generating adversarial text against real-world applications. In: Proceedings of Network and Distributed Systems Security (NDSS) Symposium (2019)","DOI":"10.14722\/ndss.2019.23138"},{"key":"10_CR18","doi-asserted-by":"crossref","unstructured":"Yeh, J., Li, S., Wu, M., Chen, W., Su, M.: Chinese word spelling correction based on N-gram ranked inverted index list. In: Proceedings of Seventh SIGHAN Workshop on Chinese Language Processing (2013)","DOI":"10.3115\/v1\/W14-6822"},{"issue":"1","key":"10_CR19","first-page":"1","volume":"20","author":"J Xiong","year":"2015","unstructured":"Xiong, J., Zhang, Q., Zhang, S., Hou, J., Cheng, X.: HANSpeller: a unified framework for Chinese spelling correction. Comput. Linguist. Chin. Lang. Process. 20(1), 1\u201322 (2015)","journal-title":"Comput. Linguist. Chin. Lang. Process."},{"key":"10_CR20","doi-asserted-by":"crossref","unstructured":"Karan, M., Snajder, J.: Cross-domain detection of abusive language online. In: Proceedings of 2nd Workshop on Abusive Language Online (2018)","DOI":"10.18653\/v1\/W18-5117"},{"key":"10_CR21","doi-asserted-by":"crossref","unstructured":"Nobata, C., Tetreault, J., Thomas, A., Mehdad, Y., Chang, Y.: Abusive language detection in online user content. In: 25th International Conference on World Wide Web (2016)","DOI":"10.1145\/2872427.2883062"},{"key":"10_CR22","unstructured":"Wikipedia, Commonly used words in Modern Chinese. https:\/\/zh.wikipedia.org\/wiki\/\u73b0\u4ee3\u6c49\u8bed\u5e38\u7528\u5b57\u8868. Accessed 07 Apr 2021"},{"key":"10_CR23","unstructured":"Li, J., et al.: TEXTSHIELD: robust text classification based on multimodal embedding and neural machine translation. In: Proceedings of 29th USENIX Security Symposium (2020)"},{"key":"10_CR24","doi-asserted-by":"crossref","unstructured":"Yu, J., Jian, X., Xin, H., Song, Y.: Joint embeddings of Chinese words, characters, and fine-grained subcharacter components. In: Proceedings of Conference on Empirical Methods in Natural Language Processing (2017)","DOI":"10.18653\/v1\/D17-1027"},{"key":"10_CR25","doi-asserted-by":"crossref","unstructured":"Cao, S., Lu, W., Zhou, J., Li, X.: cw2vec: learning Chinese word embeddings with stroke n-gram information. In: Proceedings of AAAI (2018)","DOI":"10.1609\/aaai.v32i1.12029"},{"issue":"2","key":"10_CR26","doi-asserted-by":"publisher","first-page":"299","DOI":"10.1007\/s10579-012-9197-9","volume":"47","author":"T Chen","year":"2013","unstructured":"Chen, T., Kan, M.: Creating a live, public short message service corpus: the NUS SMS corpus. Lang. Resour. Eval. 47(2), 299\u2013335 (2013). https:\/\/doi.org\/10.1007\/s10579-012-9197-9","journal-title":"Lang. Resour. Eval."},{"key":"10_CR27","unstructured":"SpamMessage. https:\/\/github.com\/hrwhisper\/SpamMessage. Accessed 07 Apr 2021"},{"key":"10_CR28","unstructured":"Jieba. https:\/\/github.com\/fxsjy\/jieba. Accessed 07 Apr 2021"},{"key":"10_CR29","unstructured":"Pycorrector. https:\/\/github.com\/shibing624\/pycorrector. Accessed 07 Apr 2021"},{"key":"10_CR30","unstructured":"Scikit-learn. https:\/\/sklearn.org. Accessed 07 Apr 2021"}],"container-title":["Communications in Computer and Information Science","Cyber Security"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/978-981-16-9229-1_10","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2023,1,24]],"date-time":"2023-01-24T01:14:37Z","timestamp":1674522877000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/978-981-16-9229-1_10"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022]]},"ISBN":["9789811692284","9789811692291"],"references-count":30,"URL":"https:\/\/doi.org\/10.1007\/978-981-16-9229-1_10","relation":{},"ISSN":["1865-0929","1865-0937"],"issn-type":[{"type":"print","value":"1865-0929"},{"type":"electronic","value":"1865-0937"}],"subject":[],"published":{"date-parts":[[2022]]},"assertion":[{"value":"21 January 2022","order":1,"name":"first_online","label":"First Online","group":{"name":"ChapterHistory","label":"Chapter History"}},{"value":"CNCERT","order":1,"name":"conference_acronym","label":"Conference Acronym","group":{"name":"ConferenceInfo","label":"Conference Information"}},{"value":"China Cyber Security Annual Conference","order":2,"name":"conference_name","label":"Conference Name","group":{"name":"ConferenceInfo","label":"Conference Information"}},{"value":"Beijing","order":3,"name":"conference_city","label":"Conference City","group":{"name":"ConferenceInfo","label":"Conference Information"}},{"value":"China","order":4,"name":"conference_country","label":"Conference Country","group":{"name":"ConferenceInfo","label":"Conference Information"}},{"value":"2021","order":5,"name":"conference_year","label":"Conference Year","group":{"name":"ConferenceInfo","label":"Conference Information"}},{"value":"20 July 2021","order":7,"name":"conference_start_date","label":"Conference Start Date","group":{"name":"ConferenceInfo","label":"Conference Information"}},{"value":"21 July 2021","order":8,"name":"conference_end_date","label":"Conference End Date","group":{"name":"ConferenceInfo","label":"Conference Information"}},{"value":"18","order":9,"name":"conference_number","label":"Conference Number","group":{"name":"ConferenceInfo","label":"Conference Information"}},{"value":"cncert2021","order":10,"name":"conference_id","label":"Conference ID","group":{"name":"ConferenceInfo","label":"Conference Information"}},{"value":"http:\/\/conf.cert.org.cn","order":11,"name":"conference_url","label":"Conference URL","group":{"name":"ConferenceInfo","label":"Conference Information"}}]}}