{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,10]],"date-time":"2026-06-10T15:19:41Z","timestamp":1781104781557,"version":"3.54.1"},"reference-count":25,"publisher":"IGI Global Scientific Publishing","issue":"3","content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2019,7,1]]},"abstract":"<p>In the field of modern information technology, how to find information quickly, accurately and comprehensively that users really needed has become the focus of research in this field. In this article, a feature selection method based on a complex network is proposed for the structure and content characteristics of large-scale web text information. The preprocessed web text is converted into a complex network. The nodes in the network correspond to the entries in the text. The edges of the network correspond to the links between the entries in the text, and the degree of nodes and the aggregation system are used. Second, the text classification method is studied from the point of view of data sampling, and a text classification method based on density statistics is proposed. This method uses not only the density information of the text feature set in the classification process, but also the use of statistical merging criteria to get the text. The difference information of each feature has a better classification effect for large text collections.<\/p>","DOI":"10.4018\/ijaci.2019070102","type":"journal-article","created":{"date-parts":[[2019,7,19]],"date-time":"2019-07-19T09:46:35Z","timestamp":1563529595000},"page":"17-32","source":"Crossref","is-referenced-by-count":18,"title":["Web Text Categorization Based on Statistical Merging Algorithm in Big Data Environment"],"prefix":"10.4018","volume":"10","author":[{"given":"Rujuan","family":"Wang","sequence":"first","affiliation":[{"name":"College of Humanities & Sciences of Northeast Normal University, Changchun, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Gang","family":"Wang","sequence":"additional","affiliation":[{"name":"Northeast Normal University, Changchun, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"2432","reference":[{"key":"IJACI.2019070102-0","doi-asserted-by":"publisher","DOI":"10.1098\/rspb.2001.1800"},{"key":"IJACI.2019070102-1","doi-asserted-by":"crossref","DOI":"10.1007\/978-3-319-60435-0","author":"N.Dey","year":"2018","journal-title":"Internet of Things and big data analytics toward next-generation intelligence"},{"key":"IJACI.2019070102-2","doi-asserted-by":"crossref","first-page":"51","DOI":"10.1007\/978-981-10-7566-7_6","article-title":"Application of TF-IDF feature for categorizing documents of online bangla web text corpus","author":"A.Dhar","year":"2018","journal-title":"Intelligent Engineering Informatics"},{"issue":"4","key":"IJACI.2019070102-3","first-page":"691","article-title":"Survey of data mining for microblogs.","volume":"51","author":"Z.Ding","year":"2014","journal-title":"Journal of Computer Research & Development"},{"key":"IJACI.2019070102-4","doi-asserted-by":"publisher","DOI":"10.1108\/00220410410560573"},{"key":"IJACI.2019070102-5","doi-asserted-by":"publisher","DOI":"10.1108\/00220410410560573"},{"key":"IJACI.2019070102-6","article-title":"Density-based statistical merging algorithm for large data sets.","author":"B. B.Liu","year":"2015","journal-title":"Journal of Software"},{"key":"IJACI.2019070102-7","doi-asserted-by":"publisher","DOI":"10.3724\/SP.J.1001.2013.04467"},{"key":"IJACI.2019070102-8","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2004.110"},{"key":"IJACI.2019070102-9","unstructured":"Peng, M., Huang, J., Zhu, J., Huang, J., Liu, J., & School, C., et al. (2015). Mass of short texts clustering and topic extraction based on frequent itemsets. Journal of Computer Research & Development."},{"issue":"1","key":"IJACI.2019070102-10","first-page":"133","article-title":"A dynamic density-based clustering algorithm appropriate to large-scale text processing.","volume":"49","author":"Z.Qiansheng","year":"2013","journal-title":"Beijing Da Xue Xue Bao. Zi Ran Ke Xue Bao"},{"issue":"8","key":"IJACI.2019070102-11","first-page":"1","article-title":"Automatic text classification using bplion-neural network and semantic word processing.","author":"N. M.Ranjan","year":"2017","journal-title":"Imaging Science Journal"},{"key":"IJACI.2019070102-12","first-page":"225","article-title":"Using the leader algorithm with support vector machines for large data sets.","author":"E.Romero","year":"2011","journal-title":"International Conference on Artificial Neural Networks"},{"issue":"5","key":"IJACI.2019070102-13","first-page":"1019","article-title":"Dynamic assembly classification algorithm for short text.","volume":"37","author":"Y.Rui","year":"2009","journal-title":"Tien Tzu Hsueh Pao"},{"key":"IJACI.2019070102-14","doi-asserted-by":"crossref","DOI":"10.4018\/978-1-5225-2483-0","author":"A.Singh","year":"2017","journal-title":"Web Semantics for Textual and Visual Information Retrieval"},{"key":"IJACI.2019070102-15","author":"P. A.Vijaya","year":"2004","journal-title":"Leaders-subleaders: an efficient hierarchical clustering algorithm for large data sets"},{"key":"IJACI.2019070102-16","doi-asserted-by":"publisher","DOI":"10.1016\/j.patrec.2009.08.008"},{"key":"IJACI.2019070102-17","doi-asserted-by":"crossref","unstructured":"Xu, L., Fu, Y., & Li, S. (2011). Web text classifier based on an improved SVM decision tree. Journal of Soochow University.","DOI":"10.1016\/j.phpro.2012.05.312"},{"issue":"1","key":"IJACI.2019070102-18","first-page":"1986","article-title":"Text classifier based on an improved SVM decision tree.","volume":"33","author":"Z.Xu","year":"2012","journal-title":"Journal of Intelligence"},{"key":"IJACI.2019070102-19","doi-asserted-by":"publisher","DOI":"10.1109\/TKDE.2010.34"},{"key":"IJACI.2019070102-20","doi-asserted-by":"publisher","DOI":"10.4028\/www.scientific.net\/AMR.945-949.2306"},{"key":"IJACI.2019070102-21","doi-asserted-by":"crossref","first-page":"179","DOI":"10.1016\/j.knosys.2013.05.013","article-title":"Projected-prototype based classifier for text categorization.","volume":"49","author":"J.Zhang","year":"2013","journal-title":"Knowledge-Based Systems"},{"issue":"6","key":"IJACI.2019070102-22","first-page":"936","article-title":"An improved knn text categorization algorithm by adopting cluster technology.","volume":"22","author":"X. F.Zhang","year":"2009","journal-title":"Pattern Recognition & Artificial Intelligence"},{"key":"IJACI.2019070102-23","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-38562-9_3"},{"issue":"6","key":"IJACI.2019070102-24","first-page":"1352","article-title":"Extraction of gene\/protein names involved in each stage of spermatogenesis based on literature mining.","volume":"51","author":"J.Zhu","year":"2014","journal-title":"Journal of Computer Research & Development"}],"container-title":["International Journal of Ambient Computing and Intelligence"],"original-title":[],"language":"ng","link":[{"URL":"https:\/\/www.igi-global.com\/viewtitle.aspx?TitleId=233816","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2022,5,6]],"date-time":"2022-05-06T14:34:21Z","timestamp":1651847661000},"score":1,"resource":{"primary":{"URL":"https:\/\/services.igi-global.com\/resolvedoi\/resolve.aspx?doi=10.4018\/IJACI.2019070102"}},"subtitle":[""],"short-title":[],"issued":{"date-parts":[[2019,7,1]]},"references-count":25,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2019,7]]}},"URL":"https:\/\/doi.org\/10.4018\/ijaci.2019070102","relation":{},"ISSN":["1941-6237","1941-6245"],"issn-type":[{"value":"1941-6237","type":"print"},{"value":"1941-6245","type":"electronic"}],"subject":[],"published":{"date-parts":[[2019,7,1]]}}}