{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,10,12]],"date-time":"2025-10-12T04:00:49Z","timestamp":1760241649012,"version":"build-2065373602"},"reference-count":26,"publisher":"MDPI AG","issue":"8","license":[{"start":{"date-parts":[[2018,7,24]],"date-time":"2018-07-24T00:00:00Z","timestamp":1532390400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"DOI":"10.13039\/501100010880","name":"STATE GRID Corporation of China","doi-asserted-by":"publisher","award":["Headquarters Technology Program 2017"],"award-info":[{"award-number":["Headquarters Technology Program 2017"]}],"id":[{"id":"10.13039\/501100010880","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Algorithms"],"abstract":"<jats:p>The exponential increase in online reviews and recommendations makes document classification and sentiment analysis a hot topic in academic and industrial research. Traditional deep learning based document classification methods require the use of full textual information to extract features. In this paper, in order to tackle long document, we proposed three methods that use local convolutional feature aggregation to implement document classification. The first proposed method randomly draws blocks of continuous words in the full document. Each block is then fed into the convolution neural network to extract features and then are concatenated together to output the classification probability through a classifier. The second model improves the first by capturing the contextual order information of the sampled blocks with a recurrent neural network. The third model is inspired by the recurrent attention model (RAM), in which a reinforcement learning module is introduced to act as a controller for selecting the next block position based on the recurrent state. Experiments on our collected four-class arXiv paper dataset show that the three proposed models all perform well, and the RAM model achieves the best test accuracy with the least information.<\/jats:p>","DOI":"10.3390\/a11080109","type":"journal-article","created":{"date-parts":[[2018,7,24]],"date-time":"2018-07-24T11:00:07Z","timestamp":1532430007000},"page":"109","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":12,"title":["Long Length Document Classification by Local Convolutional Feature Aggregation"],"prefix":"10.3390","volume":"11","author":[{"given":"Liu","family":"Liu","sequence":"first","affiliation":[{"name":"School of Electronic and Information Engineering, Nanjing University of Information Science and Technology, Nanjing 210044, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Kaile","family":"Liu","sequence":"additional","affiliation":[{"name":"State Grid Corporation of China, Beijing 100031, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Zhenghai","family":"Cong","sequence":"additional","affiliation":[{"name":"NARI Group Corporation of China\/State Grid Electric Power Research Institute, Nanjing 211106, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jiali","family":"Zhao","sequence":"additional","affiliation":[{"name":"State Grid Corporation of China, Beijing 100031, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yefei","family":"Ji","sequence":"additional","affiliation":[{"name":"NARI Group Corporation of China\/State Grid Electric Power Research Institute, Nanjing 211106, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jun","family":"He","sequence":"additional","affiliation":[{"name":"School of Electronic and Information Engineering, Nanjing University of Information Science and Technology, Nanjing 210044, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1968","published-online":{"date-parts":[[2018,7,24]]},"reference":[{"key":"ref_1","unstructured":"Wang, S., and Manning, C.D. (2012, January 8\u201314). Baselines and bigrams: Simple, good sentiment and topic classification. Proceedings of the 50th Annual Meeting of the Association for Computational Linguistics: Short Papers-Volume 2, Jeju Island, Korea."},{"key":"ref_2","doi-asserted-by":"crossref","unstructured":"Pang, B., and Lee, L. (2005, January 25\u201330). Seeing stars: Exploiting class relationships for sentiment categorization with respect to rating scales. Proceedings of the 43rd Annual Meeting on Association for Computational Linguistics, Ann Arbor, MI, USA.","DOI":"10.3115\/1219840.1219855"},{"key":"ref_3","doi-asserted-by":"crossref","unstructured":"Kim, Y. (2014). Convolutional Neural Networks for Sentence Classification. Eprint Arxiv.","DOI":"10.3115\/v1\/D14-1181"},{"key":"ref_4","doi-asserted-by":"crossref","unstructured":"Lai, S., Xu, L., Liu, K., and Zhao, J. (2015, January 25\u201330). Recurrent Convolutional Neural Networks for Text Classification. Proceedings of the Twenty-Ninth AAAI Conference on Artificial Intelligence, Austin, TX, USA.","DOI":"10.1609\/aaai.v29i1.9513"},{"key":"ref_5","doi-asserted-by":"crossref","unstructured":"Yang, Z., Yang, D., Dyer, C., He, X., Smola, A., and Hovy, E. (2016, January 12\u201317). Hierarchical attention net-works for document classification. Proceedings of the NAACL-HLT, San Diego, CA, USA.","DOI":"10.18653\/v1\/N16-1174"},{"key":"ref_6","unstructured":"Zhang, X., Zhao, J., and Lecun, Y. (2015, January 7\u201312). Character-level Convolutional Networks for Text Classification. Proceedings of the NIPS\u201915 28th International Conference on Neural Information Processing Systems, Montreal, QC, Canada."},{"key":"ref_7","doi-asserted-by":"crossref","unstructured":"Tang, D., Qin, B., and Liu, T. (2015, January 17). Document Modeling with Gated Recurrent Neural Network for Sentiment Classification. Proceedings of the Conference on Empirical Methods in Natural Language Processing, Lisboa, Portugal.","DOI":"10.18653\/v1\/D15-1167"},{"key":"ref_8","doi-asserted-by":"crossref","unstructured":"Hu, M., and Liu, B. (2004, January 22\u201325). Mining and summarizing customer reviews. Proceedings of the Tenth ACM SIGKDD International Conference on Knowledge Discovery and Data Mining, Seattle, WA, USA.","DOI":"10.1145\/1014052.1014073"},{"key":"ref_9","unstructured":"Mnih, V., Heess, N., and Graves, A. (2014, January 8\u201313). Recurrent models of visual attention. Proceedings of the Advances in Neural Information Processing Systems, Montreal, QC, Canada."},{"key":"ref_10","unstructured":"Simonyan, K., and Zisserman, A. (2014). Very deep convolutional networks for large-scale image recognition. arXiv."},{"key":"ref_11","doi-asserted-by":"crossref","unstructured":"Szegedy, C., Liu, W., Jia, Y., Sermanet, P., Reed, S., Anguelov, D., Erhan, D., Vanhoucke, V., and Rabinovich, A. (2015, January 7\u201312). Going deeper with convolutions. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Boston, MA, USA.","DOI":"10.1109\/CVPR.2015.7298594"},{"key":"ref_12","doi-asserted-by":"crossref","unstructured":"He, K., Zhang, X., Ren, S., and Sun, J. (2016, January 27\u201330). Deep Residual Learning for Image Recognition. Proceedings of the 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Las Vegas, NV, USA.","DOI":"10.1109\/CVPR.2016.90"},{"key":"ref_13","unstructured":"Maas, A.L., Daly, R.E., Pham, P.T., Huang, D., Na, A.Y., and Potts, D. (2011, January 19\u201324). Learning word vectors for sentiment analysis. Proceedings of the 49th Annual Meeting of the Association for Computational Linguistics: Human Language Technologies-Volume 1, Portland, OR, USA."},{"key":"ref_14","unstructured":"Blunsom, P., Cohen, S., Dhillon, P., and Liang, P. (2015). Relation Extraction: Perspective from Convolutional Neural Networks. Workshop on Vector Modeling for NLP, Association for Computational Linguistics."},{"key":"ref_15","doi-asserted-by":"crossref","first-page":"1735","DOI":"10.1162\/neco.1997.9.8.1735","article-title":"Long short-term memory","volume":"9","author":"Hochreiter","year":"1997","journal-title":"Neural Comput."},{"key":"ref_16","unstructured":"Sutskever, I., Vinyals, O., and Le, Q.V. (2014, January 8\u201313). Sequence to Sequence Learning with Neural Networks. Proceedings of the 27th International Conference on Neural Information Processing Systems, Montreal, QC, Canada."},{"key":"ref_17","unstructured":"Bahdanau, D., Cho, K., and Bengio, Y. (2014). Neural Machine Translation by Jointly Learning to Align and Translate. arXiv."},{"key":"ref_18","doi-asserted-by":"crossref","unstructured":"Nallapati, R., Zhou, B., dos Santos, C., Gulcehre, C., and Xiang, B. (2016). Abstractive Text Summarization Using Sequence-to-Sequence RNNs and Beyond. arXiv.","DOI":"10.18653\/v1\/K16-1028"},{"key":"ref_19","unstructured":"Wang, Z., He, W., Wu, H., Wu, H., Li, W., Wang, H., and Chen, E. (2016). Chinese Poetry Generation with Planning based Neural Network. arXiv."},{"key":"ref_20","unstructured":"Liu, P., Qiu, X., and Huang, X. (2016). Recurrent neural network for text classification with multi-task learning. arXiv."},{"key":"ref_21","doi-asserted-by":"crossref","first-page":"855","DOI":"10.1109\/TPAMI.2008.137","article-title":"A novel connectionist system for unconstrained handwriting recognition","volume":"31","author":"Graves","year":"2009","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_22","doi-asserted-by":"crossref","unstructured":"Miao, Y., Gowayyed, M., and Metze, F. (2015, January 13\u201317). EESEN: End-to-end speech recognition using deep RNN models and WFST-based decoding. Proceedings of the 2015 IEEE Workshop on Automatic Speech Recognition and Understanding (ASRU), Scottsdale, AZ, USA.","DOI":"10.1109\/ASRU.2015.7404790"},{"key":"ref_23","doi-asserted-by":"crossref","unstructured":"Pennington, J., Socher, R., and Manning, C. (2014, January 25\u201329). Glove: Global Vectors for Word Representation. Proceedings of the 27th Conference on Empirical Methods in Natural Language Processing, Doha, Qatar.","DOI":"10.3115\/v1\/D14-1162"},{"key":"ref_24","unstructured":"(2018, July 23). Tensorflow. Available online: https:\/\/github.com\/tensorflow\/tensorflow."},{"key":"ref_25","first-page":"281","article-title":"Random search for hyper-parameter optimization","volume":"13","author":"Stra","year":"2012","journal-title":"J. Mach. Learn. Res."},{"key":"ref_26","unstructured":"Snoek, J., Larochell, H., and Adams, R.P. (2012, January 3\u20136). Practical bayesian optimization of machine learning algorithms. Proceedings of the Advances in neural information processing systems, Lake Tahoe, NV, USA."}],"container-title":["Algorithms"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/1999-4893\/11\/8\/109\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,11]],"date-time":"2025-10-11T15:13:58Z","timestamp":1760195638000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/1999-4893\/11\/8\/109"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2018,7,24]]},"references-count":26,"journal-issue":{"issue":"8","published-online":{"date-parts":[[2018,8]]}},"alternative-id":["a11080109"],"URL":"https:\/\/doi.org\/10.3390\/a11080109","relation":{},"ISSN":["1999-4893"],"issn-type":[{"type":"electronic","value":"1999-4893"}],"subject":[],"published":{"date-parts":[[2018,7,24]]}}}