{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,9]],"date-time":"2026-06-09T15:57:56Z","timestamp":1781020676420,"version":"3.54.1"},"reference-count":61,"publisher":"SAGE Publications","issue":"1","license":[{"start":{"date-parts":[[2020,6,6]],"date-time":"2020-06-06T00:00:00Z","timestamp":1591401600000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/journals.sagepub.com\/page\/policies\/text-and-data-mining-license"}],"content-domain":{"domain":["journals.sagepub.com"],"crossmark-restriction":true},"short-container-title":["Journal of Intelligent &amp; Fuzzy Systems"],"published-print":{"date-parts":[[2020,7,17]]},"abstract":"<jats:p>Topic modeling for short texts is a challenging and interesting problem in the machine learning and knowledge discovery domains. Nowadays, millions of documents published on the internet from various sources. Internet websites are full of various topics and information, but there is a lot of similarity between topics, contents, and total quality of sources, which causes data repetition and gives the user the same information. Another issue is data sparsity and ambiguity because the length of the short text is limited, which causes unsatisfactory results and give irrelevant results to end-users. All these mentioned issues in short texts made an interesting topic for researchers to use machine learning and knowledge discovery techniques to discover underlying topics from a massive amount of data. In this paper, we propose a combination of deep reinforcement learning (RL) and semantics-assisted non-negative matrix factorization model to extract meaningful and underlying topics from short document contents. The main objective of this work is to reduce the problem of repetitive information and data sparsity in short texts to help the users to get meaningful and relevant contents. Furthermore, our propose model reviews an issue of the Seq2Seq approach based on the reinforcement learning perspective and provides a combination of reinforcement learning and SeaNMF formulation using the block coordinate descent algorithm. Moreover, we compare different real-world datasets by using numerical calculation and present a couple of state-of-art models to get better performance on short text document topic modeling. Based on experimental results and comparative analysis, our propose model outperforms the state of art techniques in terms of short document topic modeling.<\/jats:p>","DOI":"10.3233\/jifs-191690","type":"journal-article","created":{"date-parts":[[2020,6,9]],"date-time":"2020-06-09T13:02:08Z","timestamp":1591707728000},"page":"753-770","update-policy":"https:\/\/doi.org\/10.1177\/sage-journals-update-policy","source":"Crossref","is-referenced-by-count":15,"title":["Topic modeling in short-text using non-negative matrix factorization based on deep reinforcement learning"],"prefix":"10.1177","volume":"39","author":[{"given":"Zeinab","family":"Shahbazi","sequence":"first","affiliation":[{"name":"Department of Computer Engineering, Jeju National University, Jejusi, Jeju Special Self-Governing Provience, Korea"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yung-Cheol","family":"Byun","sequence":"additional","affiliation":[{"name":"Department of Computer Engineering, Jeju National University, Jejusi, Jeju Special Self-Governing Provience, Korea"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"179","published-online":{"date-parts":[[2020,6,6]]},"reference":[{"key":"e_1_3_2_2_2","doi-asserted-by":"crossref","unstructured":"ZhaoW.X. JiangJ. WengJ. HeJ. LimE.-P. YanH. and LiX. Comparing twitter and traditional media using topic models. In European conference on information retrieval (2011) pp. 338\u2013349. Springer.","DOI":"10.1007\/978-3-642-20161-5_34"},{"key":"e_1_3_2_3_2","doi-asserted-by":"crossref","unstructured":"SriramB. FuhryD. DemirE. FerhatosmanogluH. and DemirbasM. Short text classification in twitter to improve information filtering. In Proceedings of the 33rd international ACM SIGIR conference on Research and development in information retrieval (2010) pp. 841\u2013842. ACM.","DOI":"10.1145\/1835449.1835643"},{"key":"e_1_3_2_4_2","doi-asserted-by":"crossref","unstructured":"HongL. and DavisonB.D. Empirical study of topic modeling in twitter. In Proceedings of the first workshop on social media analytics (2010) pp. 80\u201388. acm.","DOI":"10.1145\/1964858.1964870"},{"key":"e_1_3_2_5_2","unstructured":"WangZ. WangH. Understanding short texts. 2016."},{"key":"e_1_3_2_6_2","doi-asserted-by":"crossref","unstructured":"YanX. GuoJ. LanY. and Cheng.X. A biterm topic model for short texts. In Proceedings of the 22nd international conference on World Wide Web (2013) pp. 1445\u20131456. ACM.","DOI":"10.1145\/2488388.2488514"},{"key":"e_1_3_2_7_2","first-page":"993","article-title":"Latent dirichlet allocation","volume":"3","author":"Blei D.M.","year":"2003","unstructured":"BleiD.M., NgA.Y. and JordanM.I., Latent dirichlet allocation, Journal of Machine Learning Research 3(Jan) (2003), 993\u20131022.","journal-title":"Journal of Machine Learning Research"},{"key":"e_1_3_2_8_2","doi-asserted-by":"publisher","DOI":"10.1002\/(SICI)1097-4571(199009)41:6<391::AID-ASI1>3.0.CO;2-9"},{"key":"e_1_3_2_9_2","doi-asserted-by":"publisher","DOI":"10.1145\/3130348.3130370"},{"key":"e_1_3_2_10_2","doi-asserted-by":"publisher","DOI":"10.1038\/44565"},{"key":"e_1_3_2_11_2","doi-asserted-by":"publisher","DOI":"10.1007\/s10618-014-0384-8"},{"key":"e_1_3_2_12_2","doi-asserted-by":"crossref","unstructured":"KuangD. ChooJ. and ParkH. Nonnegative matrix factorization for interactive topic modeling and document clustering In Partitional Clustering Algorithms (2015) pp. 215\u2013243. Springer.","DOI":"10.1007\/978-3-319-09259-1_7"},{"key":"e_1_3_2_13_2","doi-asserted-by":"publisher","DOI":"10.1007\/s10898-014-0247-2"},{"key":"e_1_3_2_14_2","unstructured":"QuanX. KitC. GeY. and PanS.J. Short and sparse text topicmodeling via self-aggregation. In Twenty-Fourth International Joint Conference on Artificial Intelligence 2015."},{"key":"e_1_3_2_15_2","doi-asserted-by":"crossref","unstructured":"ZuoY. WuJ. ZhangH. LinH. WangF. XuK. and XiongH. Topic modeling of short texts: A pseudodocument view. In Proceedings of the 22nd ACM SIGKDD international conference on knowledge discovery and data mining (2016) pp. 2105\u20132114. ACM.","DOI":"10.1145\/2939672.2939880"},{"key":"e_1_3_2_16_2","unstructured":"MikolovT. ChenK. CorradoG. and DeanJ. Efficient estimation of word representations in vector space. arXiv preprint arXiv:1301.3781 2013."},{"key":"e_1_3_2_17_2","doi-asserted-by":"crossref","unstructured":"PenningtonJ. SocherR. and ManningC. Glove: Global vectors for word representation. In Proceedings of the 2014 conference on empirical methods in natural language processing (EMNLP) (2014) pp. 1532\u20131543.","DOI":"10.3115\/v1\/D14-1162"},{"key":"e_1_3_2_18_2","doi-asserted-by":"crossref","unstructured":"LiC. WangH. ZhangZ. SunA. and MaZ. Topic modeling for short texts with auxiliary word embeddings. In Proceedings of the 39th International ACM SIGIR conference onResearch and Development in Information Retrieval (2016) pp. 165\u2013174. ACM.","DOI":"10.1145\/2911451.2911499"},{"key":"e_1_3_2_19_2","unstructured":"SridharV.K.R. Unsupervised topicmodeling for short texts using distributed representations of words. In Proceedings of the 1st workshop on vector space modeling for natural language processing (2015) pp. 192\u2013200."},{"key":"e_1_3_2_20_2","doi-asserted-by":"crossref","unstructured":"XunG. GopalakrishnanV. MaF. LiY. GaoJ. and ZhangA. Topic discovery for short texts using word embeddings. In 2016 IEEE 16th international conference on data mining (ICDM) (2016) pp. 1299\u20131304. IEEE.","DOI":"10.1109\/ICDM.2016.0176"},{"key":"e_1_3_2_21_2","unstructured":"MikolovT. SutskeverI. ChenK. CorradoGreg S. and DeanJeff. Distributed representations of words and phrases and their compositionality. In Advances in neural information processing systems (2013) pp. 3111\u20133119."},{"key":"e_1_3_2_22_2","unstructured":"LevyO. GoldbergY. Neural word embedding as implicit matrix factorization. In Advances in neural information processing systems (2014) pp. 2177\u20132185."},{"key":"e_1_3_2_23_2","doi-asserted-by":"publisher","DOI":"10.1162\/tacl_a_00134"},{"key":"e_1_3_2_24_2","doi-asserted-by":"crossref","unstructured":"PhanX.-H. NguyenL-M. and HoriguchiS. Learning to classify short and sparse text & web with hidden topics from large-scale data collections. In Proceedings of the 17th international conference on World Wide Web (2008) pp. 91\u2013100.","DOI":"10.1145\/1367497.1367510"},{"key":"e_1_3_2_25_2","doi-asserted-by":"crossref","unstructured":"JinO. LiuNathan N. ZhaoK. YuY. and YangQ. Transferring topical knowledge from auxiliary long texts for short text clustering. In Proceedings of the 20th ACM international conference on Information and knowledge management (2011) pp. 775\u2013784.","DOI":"10.1145\/2063576.2063689"},{"key":"e_1_3_2_26_2","doi-asserted-by":"crossref","unstructured":"WengJ. LimE-P. JiangJ. and HeQ. Twitterrank: finding topic-sensitive influential twitterers. In Proceedings of the third ACM international conference onWeb search and data mining (2010) pp. 261\u2013270.","DOI":"10.1145\/1718487.1718520"},{"key":"e_1_3_2_27_2","doi-asserted-by":"crossref","unstructured":"MehrotraR. SannerS. BuntineW. and XieL. Improving lda topic models for microblogs via tweet pooling and automatic labeling. In Proceedings of the 36th international ACM SIGIR conference on Research and development in information retrieval (2013) pp. 889\u2013892.","DOI":"10.1145\/2484028.2484166"},{"key":"e_1_3_2_28_2","doi-asserted-by":"publisher","DOI":"10.1007\/s10115-015-0882-z"},{"key":"e_1_3_2_29_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.neucom.2018.03.086"},{"key":"e_1_3_2_30_2","doi-asserted-by":"crossref","unstructured":"RamageD. HallD. NallapatiR. and ManningChristopher D. Labeled lda: A supervised topic model for credit attribution in multi-labeled corpora. In Proceedings of the 2009 Conference on Empirical Methods in Natural Language Processing: Volume 1-Volume 1 (2009) pp. 248\u2013256. Association for Computational Linguistics.","DOI":"10.3115\/1699510.1699543"},{"key":"e_1_3_2_31_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.jocs.2017.10.012"},{"key":"e_1_3_2_32_2","doi-asserted-by":"crossref","unstructured":"LinT. TianW. MeiQ. and ChengH. The dual-sparse topic model: mining focused topics and focused terms in short text. In Proceedings of the 23rd international conference on World wide web (2014) pp. 539\u2013550.","DOI":"10.1145\/2566486.2567980"},{"key":"e_1_3_2_33_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.ins.2017.02.007"},{"key":"e_1_3_2_34_2","doi-asserted-by":"publisher","DOI":"10.1109\/TVCG.2013.212"},{"key":"e_1_3_2_35_2","doi-asserted-by":"crossref","unstructured":"KimH. ChooJ. KimJ. ReddyChandan K. and ParkH. Simultaneous discovery of common and discriminative topics via joint nonnegative matrix factorization. In Proceedings of the 21th ACM SIGKDD International Conference on Knowledge Discovery and Data Mining (2015) pp. 567\u2013576. ACM.","DOI":"10.1145\/2783258.2783338"},{"key":"e_1_3_2_36_2","doi-asserted-by":"crossref","unstructured":"YanX. GuoJ. LiuS. ChengX. and WangY. Learning topics in short texts by non-negative matrix factorization on term correlation matrix. In proceedings of the 2013 SIAM International Conference on Data Mining (2013) pp. 749\u2013757. SIAM.","DOI":"10.1137\/1.9781611972832.83"},{"key":"e_1_3_2_37_2","doi-asserted-by":"crossref","unstructured":"ShiT. KangK. ChooJ. and ReddyChandan K. Short-text topic modeling via non-negative matrix factorization enriched with local word-context correlations. In Proceedings of the 2018 World Wide Web Conference (2018) pp. 1105\u20131114.","DOI":"10.1145\/3178876.3186009"},{"key":"e_1_3_2_38_2","unstructured":"LuongM.-T. PhamH. and ManningChristopher D. Effective approaches to attention-based neural machine translation. arXiv preprint arXiv:1508.04025 2015."},{"key":"e_1_3_2_39_2","unstructured":"WuY. SchusterM. ChenZ. LeQuoc V. NorouziM. MachereyW. KrikunM. CaoY. GaoQ. MachereyK. et al. Google\u2019s neural machine translation system: Bridging the gap between human and machine translation. arXiv preprint arXiv:1609.08144 2016."},{"key":"e_1_3_2_40_2","unstructured":"BahdanauD. ChoK. and BengioY. Neural machine translation by jointly learning to align and translate. arXiv preprint arXiv:1409.0473 2014."},{"key":"e_1_3_2_41_2","unstructured":"ShenS. ChengY. HeZ. HeW. WuH. SunM. and LiuY. Minimum risk training for neural machine translation. arXiv preprint arXiv:1512.02433 2015."},{"key":"e_1_3_2_42_2","doi-asserted-by":"crossref","unstructured":"RushAlexander M. ChopraS. and WestonJ. A neural attention model for abstractive sentence summarization. arXiv preprint arXiv:1509.00685 2015.","DOI":"10.18653\/v1\/D15-1044"},{"key":"e_1_3_2_43_2","doi-asserted-by":"crossref","unstructured":"ChopraS. AuliM. and RushAlexander M. Abstractive sentence summarization with attentive recurrent neural networks. In Proceedings of the 2016 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (2016) pp. 93\u201398.","DOI":"10.18653\/v1\/N16-1012"},{"key":"e_1_3_2_44_2","doi-asserted-by":"publisher","DOI":"10.1007\/s11390-017-1758-3"},{"key":"e_1_3_2_45_2","unstructured":"SeeA. LiuPeter J. and ManningChristopher D. Get to the point: Summarization with pointer-generator networks. arXiv preprint arXiv:1704.04368 2017."},{"key":"e_1_3_2_46_2","doi-asserted-by":"crossref","unstructured":"NallapatiR. ZhouB. GulcehreC. and XiangB. et al. Abstractive text summarization using sequence-tosequence rnns and beyond. arXiv preprint arXiv:1602.06023 2016.","DOI":"10.18653\/v1\/K16-1028"},{"key":"e_1_3_2_47_2","doi-asserted-by":"crossref","unstructured":"BahdanauD. ChorowskiJ. SerdyukD. BrakelP. and BengioY. End-to-end attentionbased large vocabulary speech recognition. In 2016 IEEE international conference on acoustics speech and signal processing (ICASSP) (2016) pp. 4945\u20134949. IEEE.","DOI":"10.1109\/ICASSP.2016.7472618"},{"key":"e_1_3_2_48_2","unstructured":"GravesA. and JaitlyN. Towards end-to-end speech recognition with recurrent neural networks. In International conference on machine learning (2014) pp. 1764\u20131772."},{"key":"e_1_3_2_49_2","doi-asserted-by":"crossref","unstructured":"MiaoY. GowayyedM. and MetzeF. Eesen: End-to-end speech recognition using deep rnn models and wfst-based decoding. In 2015 IEEE Workshop on Automatic Speech Recognition and Understanding (ASRU) (2015) pp. 167\u2013174. IEEE.","DOI":"10.1109\/ASRU.2015.7404790"},{"key":"e_1_3_2_50_2","unstructured":"BengioS. VinyalsO. JaitlyN. and ShazeerN. Scheduled sampling for sequence prediction with recurrent neural networks. In Advances in Neural Information Processing Systems (2015) pp. 1171\u20131179."},{"key":"e_1_3_2_51_2","doi-asserted-by":"publisher","DOI":"10.4304\/tpls.2.1.43-53"},{"key":"e_1_3_2_52_2","first-page":"813","article-title":"Semi-supervised sequence modeling with syntactic topic models","volume":"5","author":"Li W.","year":"2005","unstructured":"LiW. and McCallumA., Semi-supervised sequence modeling with syntactic topic models, In AAAI 5 (2005), pp. 813\u2013818.","journal-title":"AAAI"},{"key":"e_1_3_2_53_2","unstructured":"SutskeverI. VinyalsO. and LeQuoc V. Sequence to sequence learning with neural networks. In Advances in neural information processing systems (2014) pp. 3104\u20133112."},{"key":"e_1_3_2_54_2","doi-asserted-by":"publisher","DOI":"10.21248\/jlcl.27.2012.158"},{"key":"e_1_3_2_55_2","unstructured":"RiedlM. and BiemannC. How text segmentation algorithms gain from topic models. In Proceedings of the 2012 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (2012) pp. 553\u2013557. Association for Computational Linguistics."},{"key":"e_1_3_2_56_2","unstructured":"FisherD. KozdobaM. and MannorS. Topic modeling via full dependence mixtures. arXiv preprint arXiv:1906.06181 2019."},{"key":"e_1_3_2_57_2","doi-asserted-by":"crossref","unstructured":"VedantamR. Lawrence ZitnickC. and ParikhD. Cider: Consensus-based image description evaluation. In Proceedings of the IEEE conference on computer vision and pattern recognition (2015) pp. 4566\u20134575.","DOI":"10.1109\/CVPR.2015.7299087"},{"key":"e_1_3_2_58_2","doi-asserted-by":"crossref","unstructured":"PapineniK. RoukosS. WardT. and ZhuW.-J. Bleu: a method for automatic evaluation of machine translation. In Proceedings of the 40th annual meeting on association for computational linguistics (2002) pp. 311\u2013318. Association for Computational Linguistics.","DOI":"10.3115\/1073083.1073135"},{"key":"e_1_3_2_59_2","unstructured":"LinC.-Y. Looking for a few good metrics: Automatic summarization evaluation-how many samples are enough? In NTCIR 2004."},{"key":"e_1_3_2_60_2","unstructured":"BanerjeeS. and LavieA. Meteor: An automatic metric for mt evaluation with improved correlation with human judgments. In Proceedings of the acl workshop on intrinsic and extrinsic evaluation measures for machine translation and\/or summarization (2005) pp. 65\u201372."},{"key":"e_1_3_2_61_2","doi-asserted-by":"crossref","unstructured":"ZubiagaA. and JiH. Harnessing web page directories for large-scale classification of tweets. In Proceedings of the 22nd international conference on world wide web (2013) pp. 225\u2013226. ACM.","DOI":"10.1145\/2487788.2487904"},{"key":"e_1_3_2_62_2","unstructured":"ekR. and SojkaP. Gensim\u2014statistical semantics in python. Retrieved from genism. org 2011."}],"container-title":["Journal of Intelligent &amp; Fuzzy Systems"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/journals.sagepub.com\/doi\/pdf\/10.3233\/JIFS-191690","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/journals.sagepub.com\/doi\/full-xml\/10.3233\/JIFS-191690","content-type":"application\/xml","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/journals.sagepub.com\/doi\/pdf\/10.3233\/JIFS-191690","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,4,29]],"date-time":"2026-04-29T09:41:57Z","timestamp":1777455717000},"score":1,"resource":{"primary":{"URL":"https:\/\/journals.sagepub.com\/doi\/10.3233\/JIFS-191690"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,6,6]]},"references-count":61,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2020,7,17]]}},"alternative-id":["10.3233\/JIFS-191690"],"URL":"https:\/\/doi.org\/10.3233\/jifs-191690","relation":{},"ISSN":["1064-1246","1875-8967"],"issn-type":[{"value":"1064-1246","type":"print"},{"value":"1875-8967","type":"electronic"}],"subject":[],"published":{"date-parts":[[2020,6,6]]}}}