{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T04:23:16Z","timestamp":1750220596094,"version":"3.41.0"},"reference-count":95,"publisher":"Association for Computing Machinery (ACM)","issue":"6","license":[{"start":{"date-parts":[[2020,9,28]],"date-time":"2020-09-28T00:00:00Z","timestamp":1601251200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Knowl. Discov. Data"],"published-print":{"date-parts":[[2020,12,31]]},"abstract":"<jats:p>Burstiness and overdispersion phenomena of count vectors pose significant challenges in modeling such data accurately. While the dependency assumption of the multinomial distribution causes its failure to model frequency vectors in several machine learning and data mining applications, researchers found that by extending the multinomial distribution to the Dirichlet Compound multinomial (DCM), both phenomena modeling can be addressed. However, Dirichlet distribution is not the best choice, as a prior, given its negative-correlation and equal-confidence requirements. Thus, we propose to use a flexible generalization of the Dirichlet distribution, namely, the shifted-scaled Dirichlet, as a prior to the multinomial, which grants the model a capability to better fit real data, and we call the new model the Multinomial Shifted-Scaled Dirichlet (MSSD). Given that the likelihood function plays a key role in statistical inference, e.g., in maximum likelihood estimation and Fisher information matrix investigation, we propose to improve the efficiency of computing the MSSD log-likelihood by approximating its function based on Bernoulli polynomials where the log-likelihood function is computed using the proposed mesh algorithm. Moreover, given the sparsity and high-dimensionality nature of count vectors, we propose to improve its computation efficiency by approximating the novel MSSD as a member of the exponential family of distribution, which we call EMSSD. The clustering is based on mixture models, and for learning a model, selection approach is seamlessly integrated with the estimation of the parameters. The merits of the proposed approach are validated via challenging real-world applications such as hate speech detection in Twitter, real-time recognition of criminal action, and anomaly detection in crowded scenes. Results reveal that the proposed clustering frameworks offer a good compromise between other state-of-the-art techniques and outperform other approaches previously used for frequency vectors modeling. Besides, comparing to the MSSD, the approximation EMSSD has reduced the computational complexity in high-dimensional feature spaces.<\/jats:p>","DOI":"10.1145\/3406242","type":"journal-article","created":{"date-parts":[[2020,9,29]],"date-time":"2020-09-29T04:10:30Z","timestamp":1601352630000},"page":"1-35","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":4,"title":["Probabilistic Modeling for Frequency Vectors Using a Flexible Shifted-Scaled Dirichlet Distribution Prior"],"prefix":"10.1145","volume":"14","author":[{"ORCID":"https:\/\/orcid.org\/0000-0001-9328-9218","authenticated-orcid":false,"given":"Nuha","family":"Zamzami","sequence":"first","affiliation":[{"name":"Concordia Institute for Information Systems Engineering and Department of Computer Science and Artificial Intelligence, University of Jeddah, Jeddah, Saudi Arabia"}]},{"given":"Nizar","family":"Bouguila","sequence":"additional","affiliation":[{"name":"Concordia Institute for Information Systems Engineering"}]}],"member":"320","published-online":{"date-parts":[[2020,9,28]]},"reference":[{"key":"e_1_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-540-88688-4_1"},{"key":"e_1_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1109\/AI4I.2018.8665686"},{"key":"e_1_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICMLA.2018.00112"},{"key":"e_1_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1145\/3041021.3054223"},{"key":"e_1_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1137\/1.9781611972771.40"},{"key":"e_1_2_1_6_1","first-page":"1705","article-title":"Clustering with Bregman divergences","author":"Banerjee A.","year":"2005","journal-title":"Journal of Machine Learning Research 6"},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1023\/A:1008928315401"},{"volume-title":"Smith","year":"2009","author":"Bernardo Jos\u00e9 M.","key":"e_1_2_1_8_1"},{"key":"e_1_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1109\/TKDE.2007.190726"},{"key":"e_1_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/TNN.2010.2091428"},{"volume-title":"Deriving kernels from generalized Dirichlet mixture models and applications. Information Processing 8 Management 49, 1","year":"2013","author":"Bouguila Nizar","key":"e_1_2_1_11_1"},{"key":"e_1_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.engappai.2006.01.012"},{"volume-title":"Fundamentals of statistical exponential families: With applications in statistical decision theory","series-title":"Lecture Notes-Monograph Series","author":"Brown Lawrence D.","key":"e_1_2_1_13_1"},{"volume-title":"Complex Analysis","author":"Busam Rolf","key":"e_1_2_1_14_1","doi-asserted-by":"crossref","DOI":"10.1007\/978-3-540-93983-2"},{"key":"e_1_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1145\/2396761.2396860"},{"key":"e_1_2_1_16_1","unstructured":"George Casella and R. Berger. 2002. Duxbury advanced series in statistics and decision sciences. In Statistical Inference. Thomson Learning Pacific Grove CA.  George Casella and R. Berger. 2002. Duxbury advanced series in statistics and decision sciences. In Statistical Inference. Thomson Learning Pacific Grove CA."},{"key":"e_1_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1198\/106186001317243403"},{"volume-title":"Moreno","year":"2004","author":"Chan Antoni B.","key":"e_1_2_1_18_1"},{"key":"e_1_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1017\/S1351324900000139"},{"key":"e_1_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.patcog.2012.11.021"},{"key":"e_1_2_1_21_1","volume-title":"Proceedings of the Workshop on Statistical Learning in Computer Vision (ECCV\u201904)","volume":"1","author":"Csurka Gabriella","year":"2004"},{"key":"e_1_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1109\/SSCI44817.2019.9003076"},{"volume-title":"Mixture Models and Applications","author":"Daghyani Masoud","key":"e_1_2_1_23_1"},{"volume-title":"Probability for Statistics and Machine Learning","author":"DasGupta A.","key":"e_1_2_1_24_1","doi-asserted-by":"crossref","DOI":"10.1007\/978-1-4419-9634-3"},{"volume-title":"Proceedings of the 11th International AAAI Conference on Web and Social Media.","year":"2017","author":"Davidson Thomas","key":"e_1_2_1_25_1"},{"key":"e_1_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1145\/1143844.1143881"},{"key":"e_1_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/34.990138"},{"key":"e_1_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.ins.2017.12.030"},{"key":"e_1_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/W17-3013"},{"key":"e_1_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.2307\/2982840"},{"key":"e_1_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2005.239"},{"key":"e_1_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.18637\/jss.v033.i11"},{"key":"e_1_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.2307\/2529950"},{"key":"e_1_2_1_34_1","doi-asserted-by":"publisher","DOI":"10.1145\/345508.345582"},{"key":"e_1_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2017.7952460"},{"volume-title":"Proceedings of the Advances in Neural Information Processing Systems. 487--493","year":"1999","author":"Jaakkola Tommi","key":"e_1_2_1_36_1"},{"key":"e_1_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2003.1238352"},{"volume-title":"Learning Theory and Kernel Machines","author":"Jebara Tony","key":"e_1_2_1_38_1"},{"key":"e_1_2_1_39_1","first-page":"819","article-title":"Probability product kernels","author":"Jebara Tony","year":"2004","journal-title":"Journal of Machine Learning Research 5"},{"key":"e_1_2_1_40_1","doi-asserted-by":"publisher","DOI":"10.1109\/TCOM.1967.1089532"},{"volume-title":"15th International Conference on Natural Language Processing. 155\u2013160","year":"2018","author":"Kamble Satyajit","key":"e_1_2_1_41_1"},{"key":"e_1_2_1_42_1","doi-asserted-by":"publisher","DOI":"10.1017\/S1351324996001246"},{"key":"e_1_2_1_43_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2009.5206569"},{"key":"e_1_2_1_44_1","doi-asserted-by":"publisher","DOI":"10.5244\/C.19.63"},{"key":"e_1_2_1_45_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2011.6126543"},{"key":"e_1_2_1_46_1","first-page":"1","article-title":"A survey of deep learning-based network anomaly detection","volume":"22","author":"Kwon Donghwoon","year":"2017","journal-title":"Cluster Computing"},{"key":"e_1_2_1_47_1","doi-asserted-by":"publisher","DOI":"10.1111\/j.1751-5823.2001.tb00455.x"},{"key":"e_1_2_1_48_1","unstructured":"Dan Li Dacheng Chen Jonathan Goh and See-Kiong Ng. 2018. Anomaly detection with generative adversarial networks for multivariate time series. In Proceedings of the 7th International Workshop on Big Data Streams and Heterogeneous Source Mining: Algorithms Systems Programming Models and Applications on the ACM Knowledge Discovery and Data Mining Conference August 2018 London United Kingdom.  Dan Li Dacheng Chen Jonathan Goh and See-Kiong Ng. 2018. Anomaly detection with generative adversarial networks for multivariate time series. In Proceedings of the 7th International Workshop on Big Data Streams and Heterogeneous Source Mining: Algorithms Systems Programming Models and Applications on the ACM Knowledge Discovery and Data Mining Conference August 2018 London United Kingdom."},{"key":"e_1_2_1_49_1","doi-asserted-by":"publisher","DOI":"10.1109\/18.61115"},{"key":"e_1_2_1_50_1","doi-asserted-by":"publisher","DOI":"10.1023\/B:VISI.0000029664.99615.94"},{"key":"e_1_2_1_51_1","doi-asserted-by":"publisher","DOI":"10.1145\/1102351.1102420"},{"key":"e_1_2_1_52_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2010.5539872"},{"volume-title":"Proceedings of the 17th Conference on Uncertainty in Artificial Intelligence. Morgan Kaufmann Publishers Inc., 346--353","year":"2001","author":"Margaritis Diraitris","key":"e_1_2_1_53_1"},{"key":"e_1_2_1_54_1","doi-asserted-by":"publisher","DOI":"10.1049\/iet-cvi.2016.0055"},{"key":"e_1_2_1_55_1","volume-title":"Proceedings of the AAAI-98 Workshop on Learning for Text Categorization","volume":"752","author":"McCallum Andrew","year":"1998"},{"key":"e_1_2_1_56_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2009.5206641"},{"key":"e_1_2_1_58_1","doi-asserted-by":"publisher","DOI":"10.1117\/1.JEI.27.6.063016"},{"volume-title":"Proceedings of the 4th International Workshop on Compositional Data Analysis. Universitat de Girona","year":"2011","author":"Monti G. S.","key":"e_1_2_1_59_1"},{"volume-title":"Compositional Data Analysis: Theory and Applications","author":"Monti Gianna Serafina","key":"e_1_2_1_60_1"},{"volume-title":"Proceedings of the Advances in Neural Information Processing Systems. 1385--1392","year":"2004","author":"Moreno Pedro J.","key":"e_1_2_1_61_1"},{"key":"e_1_2_1_62_1","doi-asserted-by":"publisher","DOI":"10.2307\/2333468"},{"volume-title":"Morel","year":"2005","author":"Neerchal Nagaraj K.","key":"e_1_2_1_63_1"},{"volume-title":"Dirichlet and Related Distributions: Theory, Methods and Applications","author":"Ng Kai Wang","key":"e_1_2_1_64_1"},{"key":"e_1_2_1_65_1","doi-asserted-by":"publisher","DOI":"10.5555\/2044575.2044624"},{"key":"e_1_2_1_66_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/W17-3006"},{"key":"e_1_2_1_67_1","doi-asserted-by":"publisher","DOI":"10.1111\/1467-9574.00094"},{"volume-title":"Proceedings of the 4th Berkeley Symposium on Mathematical Statistics and Probability","year":"1961","author":"R\u00e9nyi Alfr\u00e9d","key":"e_1_2_1_68_1"},{"key":"e_1_2_1_69_1","doi-asserted-by":"publisher","DOI":"10.2307\/1968409"},{"key":"e_1_2_1_70_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-17711-8_28"},{"volume":"1","volume-title":"Proceedings of the 2009 IEEE 12th International Conference on Computer Vision.","author":"Michael","key":"e_1_2_1_71_1"},{"key":"e_1_2_1_73_1","doi-asserted-by":"publisher","DOI":"10.1002\/j.1538-7305.1948.tb01338.x"},{"volume-title":"Makov","year":"1985","author":"Titterington D. Michael","key":"e_1_2_1_74_1"},{"key":"e_1_2_1_75_1","doi-asserted-by":"publisher","DOI":"10.1109\/IWCMC.2019.8766353"},{"key":"e_1_2_1_76_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-540-24672-5_34"},{"key":"e_1_2_1_77_1","doi-asserted-by":"publisher","DOI":"10.1007\/3-540-53504-7_63"},{"volume-title":"Statistical and Inductive Inference by Minimum Message Length","author":"Wallace Christopher S.","key":"e_1_2_1_78_1"},{"key":"e_1_2_1_79_1","doi-asserted-by":"publisher","DOI":"10.1023\/A:1008992619036"},{"key":"e_1_2_1_80_1","doi-asserted-by":"publisher","DOI":"10.1111\/j.2517-6161.1987.tb01695.x"},{"key":"e_1_2_1_81_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2013.441"},{"key":"e_1_2_1_82_1","doi-asserted-by":"publisher","DOI":"10.1007\/s00138-016-0746-x"},{"volume-title":"A Course of Modern Analysis","author":"Whittaker Edmund Taylor","key":"e_1_2_1_83_1"},{"key":"e_1_2_1_84_1","doi-asserted-by":"publisher","DOI":"10.1080\/10170660509509290"},{"key":"e_1_2_1_85_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10618-008-0101-6"},{"key":"e_1_2_1_86_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10618-012-0296-4"},{"key":"e_1_2_1_87_1","first-page":"27","article-title":"On recursive estimation in incomplete data models. Statistics","volume":"34","author":"Yao Jian-Feng","year":"2000","journal-title":"A Journal of Theoretical and Applied Statistics"},{"key":"e_1_2_1_88_1","doi-asserted-by":"publisher","DOI":"10.1093\/bioinformatics\/btu079"},{"volume-title":"Proceedings of the European Conference on Computer Vision. Springer, 168--180","year":"2010","author":"Yuan Fei","key":"e_1_2_1_89_1"},{"key":"e_1_2_1_90_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISIE.2019.8781154"},{"key":"e_1_2_1_91_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-92058-0_7"},{"key":"e_1_2_1_92_1","doi-asserted-by":"publisher","DOI":"10.1109\/GlobalSIP45357.2019.8969324"},{"key":"e_1_2_1_93_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10489-019-01437-0"},{"key":"e_1_2_1_94_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10489-018-1333-9"},{"key":"e_1_2_1_95_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.patcog.2019.05.038"},{"volume-title":"Bouguila N., Fan W. (eds) Mixture Models and Applications. Unsupervised and Semi-Supervised Learning","author":"Zamzami Nuha","key":"e_1_2_1_96_1"},{"key":"e_1_2_1_97_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2007.383076"}],"container-title":["ACM Transactions on Knowledge Discovery from Data"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3406242","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3406242","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T21:31:52Z","timestamp":1750195912000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3406242"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,9,28]]},"references-count":95,"journal-issue":{"issue":"6","published-print":{"date-parts":[[2020,12,31]]}},"alternative-id":["10.1145\/3406242"],"URL":"https:\/\/doi.org\/10.1145\/3406242","relation":{},"ISSN":["1556-4681","1556-472X"],"issn-type":[{"type":"print","value":"1556-4681"},{"type":"electronic","value":"1556-472X"}],"subject":[],"published":{"date-parts":[[2020,9,28]]},"assertion":[{"value":"2020-01-01","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2020-06-01","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2020-09-28","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}