{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,10]],"date-time":"2026-07-10T06:05:26Z","timestamp":1783663526081,"version":"3.55.0"},"reference-count":124,"publisher":"Springer Science and Business Media LLC","issue":"2","license":[{"start":{"date-parts":[[2025,2,3]],"date-time":"2025-02-03T00:00:00Z","timestamp":1738540800000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2025,2,3]],"date-time":"2025-02-03T00:00:00Z","timestamp":1738540800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"DOI":"10.13039\/501100002322","name":"Coordena\u00e7\u00e3o de Aperfei\u00e7oamento de Pessoal de N\u00edvel Superior","doi-asserted-by":"publisher","award":["PROEX-10712476\/D"],"award-info":[{"award-number":["PROEX-10712476\/D"]}],"id":[{"id":"10.13039\/501100002322","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100002322","name":"Coordena\u00e7\u00e3o de Aperfei\u00e7oamento de Pessoal de N\u00edvel Superior","doi-asserted-by":"publisher","award":["PROEX-10712476\/D"],"award-info":[{"award-number":["PROEX-10712476\/D"]}],"id":[{"id":"10.13039\/501100002322","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100016378","name":"Technische Universit\u00e4t Dortmund","doi-asserted-by":"crossref","id":[{"id":"10.13039\/501100016378","id-type":"DOI","asserted-by":"crossref"}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Data Min Knowl Disc"],"published-print":{"date-parts":[[2025,3]]},"abstract":"<jats:title>Abstract<\/jats:title>\n          <jats:p>We perform an extensive experimental evaluation of clustering-based outlier detection methods. These methods offer benefits such as efficiency, the possibility to capitalize on more mature evaluation measures, more developed subspace analysis for high-dimensional data and better explainability, and yet they have so-far been neglected in literature. To our knowledge, our work is the first effort to analytically and empirically study their advantages and disadvantages. Our main goal is to evaluate whether or not clustering-based techniques can compete in efficiency and effectiveness against the most studied state-of-the-art algorithms in the literature. We consider the quality of the results, the resilience against different types of data and variations in parameter configuration, the scalability, and the ability to filter out inappropriate parameter values automatically based on internal measures of clustering quality. It has been recently shown that several classic, simple, unsupervised methods surpass many deep learning approaches and, hence, remain at the state-of-the-art of outlier detection. We therefore study 14 of the best classic unsupervised methods, in particular 11 clustering-based methods and 3 non-clustering-based ones, using a consistent parameterization heuristic to identify the pros and cons of each approach. We consider 46 real and synthetic datasets with up to 125k points and 1.5k dimensions aiming to achieve plausibility with the broadest possible diversity of real-world use cases. Our results indicate that the clustering-based methods are on par with (if not surpass) the non-clustering-based ones, and we argue that clustering-based methods like KMeans\u2212\u2212 should be included as baselines in future benchmarking studies, as they often offer a competitive quality at a relatively low run time, besides several other benefits.<\/jats:p>","DOI":"10.1007\/s10618-024-01086-z","type":"journal-article","created":{"date-parts":[[2025,2,3]],"date-time":"2025-02-03T09:30:15Z","timestamp":1738575015000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":24,"title":["A comparative evaluation of clustering-based outlier detection"],"prefix":"10.1007","volume":"39","author":[{"given":"Braulio V.","family":"S\u00e1nchez Vinces","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Erich","family":"Schubert","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Arthur","family":"Zimek","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Robson L. F.","family":"Cordeiro","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2025,2,3]]},"reference":[{"key":"1086_CR1","doi-asserted-by":"publisher","unstructured":"Abe N, Zadrozny B, Langford J (2006) Outlier detection by active learning. In Proceedings of the Twelfth ACM SIGKDD International Conference on Knowledge Discovery and Data Mining, pages 504\u2013509. https:\/\/doi.org\/10.1145\/1150402.1150459","DOI":"10.1145\/1150402.1150459"},{"key":"1086_CR2","doi-asserted-by":"publisher","unstructured":"Achtert E, B\u00f6hm C, David J, Kr\u00f6ger P, Zimek A Robust Clustering in Arbitrarily Oriented Subspaces, pages 763\u2013774. https:\/\/doi.org\/10.1137\/1.9781611972788.69","DOI":"10.1137\/1.9781611972788.69"},{"key":"1086_CR3","doi-asserted-by":"publisher","first-page":"152","DOI":"10.1007\/978-3-540-71703-4_15","volume-title":"Advances in Databases: Concepts, Systems and Applications","author":"Elke Achtert","year":"2007","unstructured":"Achtert Elke, B\u00f6hm Christian, Kriegel Hans-Peter, Kr\u00f6ger Peer, M\u00fcller-Gorman Ina, Zimek Arthur (2007) Detection and Visualization of Subspace Cluster Hierarchies. In: Kotagiri Ramamohanarao, Krishna P. Radha, Mohania Mukesh, Nantajeewarawat Ekawit (eds) Advances in Databases: Concepts, Systems and Applications. Springer Berlin Heidelberg, Berlin, Heidelberg, pp 152\u2013163. https:\/\/doi.org\/10.1007\/978-3-540-71703-4_15"},{"key":"1086_CR4","doi-asserted-by":"publisher","first-page":"152","DOI":"10.1007\/978-3-540-71703-4_15","volume-title":"Advances in Databases: Concepts, Systems and Applications","author":"Elke Achtert","year":"2007","unstructured":"Achtert Elke, B\u00f6hm Christian, Kriegel Hans-Peter, Kr\u00f6ger Peer, M\u00fcller-Gorman Ina, Zimek Arthur (2007) Detection and Visualization of Subspace Cluster Hierarchies. In: Kotagiri Ramamohanarao, Krishna P. Radha, Mohania Mukesh, Nantajeewarawat Ekawit (eds) Advances in Databases: Concepts, Systems and Applications. Springer Berlin Heidelberg, Berlin, Heidelberg, pp 152\u2013163. https:\/\/doi.org\/10.1007\/978-3-540-71703-4_15"},{"key":"1086_CR5","doi-asserted-by":"publisher","first-page":"231","DOI":"10.1201\/9781315373515-10","volume-title":"Data Clustering: Algorithms and Applications","author":"Charu C. Aggarwal","year":"2018","unstructured":"Aggarwal Charu C. (2018) A Survey of Stream Clustering Algorithms. In: Aggarwal Charu C., Reddy Chandan K. (eds) Data Clustering: Algorithms and Applications. Chapman and Hall\/CRC, pp 231\u2013258. https:\/\/doi.org\/10.1201\/9781315373515-10"},{"key":"1086_CR6","doi-asserted-by":"publisher","first-page":"271","DOI":"10.1007\/978-3-031-24628-9_13","volume-title":"Machine Learning for Data Science Handbook: Data Mining and Knowledge Discovery Handbook","author":"Charu C. Aggarwal","year":"2023","unstructured":"Aggarwal Charu C. (2023) Clustering in Streams. In: Rokach Lior, Maimon Oded, Shmueli Erez (eds) Machine Learning for Data Science Handbook: Data Mining and Knowledge Discovery Handbook. Springer International Publishing, Cham, pp 271\u2013300. https:\/\/doi.org\/10.1007\/978-3-031-24628-9_13"},{"key":"1086_CR7","doi-asserted-by":"publisher","unstructured":"Aggarwal CC, Yu PS (2000) Finding generalized projected clusters in high dimensional spaces. In: Proceedings of the 2000 ACM SIGMOD International Conference on Management of Data, SIGMOD \u201900, page 70-81, New York, NY, USA. Association for Computing Machinery. ISBN 1581132174.https:\/\/doi.org\/10.1145\/342009.335383","DOI":"10.1145\/342009.335383"},{"key":"1086_CR8","doi-asserted-by":"publisher","unstructured":"Agrawal R, Gehrke J, Gunopulos D, Raghavan P (1998) Automatic subspace clustering of high dimensional data for data mining applications. In: Proceedings of the 1998 ACM SIGMOD International Conference on Management of Data, SIGMOD \u201998, page 94-105, New York, NY, USA. Association for Computing Machinery. ISBN 0897919955. https:\/\/doi.org\/10.1145\/276304.276314","DOI":"10.1145\/276304.276314"},{"key":"1086_CR9","doi-asserted-by":"crossref","unstructured":"Amer M, Goldstein M, Abdennadher S (2013) Enhancing one-class support vector machines for unsupervised anomaly detection. In: Proceedings of the ACM SIGKDD workshop on outlier detection and description, pages 8\u201315","DOI":"10.1145\/2500853.2500857"},{"key":"1086_CR10","doi-asserted-by":"crossref","unstructured":"Angiulli F, Pizzuti C (2002) Fast outlier detection in high dimensional spaces. In: European conference on principles of data mining and knowledge discovery, pages 15\u201327. Springer","DOI":"10.1007\/3-540-45681-3_2"},{"issue":"2","key":"1086_CR11","doi-asserted-by":"publisher","first-page":"145","DOI":"10.1109\/TKDE.2006.29","volume":"18","author":"F Angiulli","year":"2005","unstructured":"Angiulli F, Basta S, Pizzuti C (2005) Distance-based detection and prediction of outliers. IEEE Trans Knowl Data Eng 18(2):145\u2013160","journal-title":"IEEE Trans Knowl Data Eng"},{"issue":"2","key":"1086_CR12","doi-asserted-by":"publisher","first-page":"49","DOI":"10.1145\/304181.304187","volume":"28","author":"M Ankerst","year":"1999","unstructured":"Ankerst M, Breunig MM, Kriegel H-P, Sander J (1999) Optics: ordering points to identify the clustering structure. ACM SIGMOD Rec 28(2):49\u201360","journal-title":"ACM SIGMOD Rec"},{"issue":"1","key":"1086_CR13","doi-asserted-by":"publisher","first-page":"243","DOI":"10.1016\/j.patcog.2012.07.021","volume":"46","author":"O Arbelaitz","year":"2013","unstructured":"Arbelaitz O, Gurrutxaga I, Muguerza J, P\u00e9rez JM, Perona I (2013) An extensive comparative study of cluster validity indices. Pattern Recogn 46(1):243\u2013256","journal-title":"Pattern Recogn"},{"key":"1086_CR14","doi-asserted-by":"crossref","unstructured":"Banfield, JD, Raftery AE (1993) Model-based gaussian and non-gaussian clustering. Biometrics, pages 803\u2013821","DOI":"10.2307\/2532201"},{"key":"1086_CR15","volume-title":"Outliers in statistical data","author":"V Barnett","year":"1994","unstructured":"Barnett V, Lewis T et al (1994) Outliers in statistical data, vol 3. Wiley, New York"},{"key":"1086_CR16","doi-asserted-by":"publisher","unstructured":"Batool F, Hennig C (2021) Clustering with the Average Silhouette Width. Computational Statistics and Data Analysis, 158:107190. ISSN 01679473. https:\/\/doi.org\/10.1016\/j.csda.2021.107190","DOI":"10.1016\/j.csda.2021.107190"},{"issue":"1","key":"1086_CR17","first-page":"152","volume":"17","author":"A Benavoli","year":"2016","unstructured":"Benavoli A, Corani G, Mangili F (2016) Should we really use post-hoc tests based on mean-ranks? J Mach Learn Res 17(1):152\u2013161","journal-title":"J Mach Learn Res"},{"key":"1086_CR18","unstructured":"Bishop CM (2007) Pattern recognition and machine learning. 5th edition. ISBN 9780387310732"},{"issue":"3","key":"1086_CR19","doi-asserted-by":"publisher","first-page":"551","DOI":"10.1145\/3381028","volume":"53","author":"A Boukerche","year":"2021","unstructured":"Boukerche A, Zheng L, Alfandi O (2021) Outlier detection: methods, models, and classification. ACM Comput Surv 53(3):551\u20135537. https:\/\/doi.org\/10.1145\/3381028","journal-title":"ACM Comput Surv"},{"key":"1086_CR20","doi-asserted-by":"publisher","first-page":"262","DOI":"10.1007\/978-3-540-48247-5_28","volume-title":"Principles of Data Mining and Knowledge Discovery","author":"Markus M. Breunig","year":"1999","unstructured":"Breunig Markus M., Kriegel Hans-Peter, Ng Raymond T., Sander J\u00f6rg (1999) OPTICS-OF: Identifying Local Outliers. In: \u017bytkow Jan M., Rauch Jan (eds) Principles of Data Mining and Knowledge Discovery. Springer Berlin Heidelberg, Berlin, Heidelberg, pp 262\u2013270. https:\/\/doi.org\/10.1007\/978-3-540-48247-5_28"},{"issue":"2","key":"1086_CR21","doi-asserted-by":"publisher","first-page":"93","DOI":"10.1145\/335191.335388","volume":"29","author":"Markus M. Breunig","year":"2000","unstructured":"Breunig Markus M., Kriegel Hans-Peter, Ng Raymond T., Sander J\u00f6rg (2000) LOF: identifying density-based local outliers. ACM SIGMOD Record 29(2):93\u2013104. https:\/\/doi.org\/10.1145\/335191.335388","journal-title":"ACM SIGMOD Record"},{"issue":"1","key":"1086_CR22","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/2733381","volume":"10","author":"RJGB Campello","year":"2015","unstructured":"Campello RJGB, Moulavi D, Zimek A, Sander J (2015) Hierarchical density estimates for data clustering, visualization, and outlier detection. ACM Trans Knowl Discov Data 10(1):1\u201351. https:\/\/doi.org\/10.1145\/2733381","journal-title":"ACM Trans Knowl Discov Data"},{"issue":"2","key":"1086_CR23","doi-asserted-by":"publisher","first-page":"e1343","DOI":"10.1002\/widm.1343","volume":"10","author":"Ricardo J. G. B. Campello","year":"2020","unstructured":"Campello Ricardo J. G. B., Kr\u00f6ger Peer, Sander J\u00f6rg, Zimek Arthur (2020) Density based clustering. WIREs Data Mining Knowl Discov 10(2):e1343. https:\/\/doi.org\/10.1002\/widm.1343","journal-title":"WIREs Data Mining Knowl Discov"},{"issue":"4","key":"1086_CR24","doi-asserted-by":"publisher","first-page":"891","DOI":"10.1007\/s10618-015-0444-8","volume":"30","author":"Guilherme O. Campos","year":"2016","unstructured":"Campos Guilherme O., Zimek Arthur, Sander J\u00f6rg, Campello Ricardo J. G. B., Micenkov\u00e1 Barbora, Schubert Erich, Assent Ira, Houle Michael E. (2016) On the evaluation of unsupervised outlier detection: measures, datasets, and an empirical study. Data Mining Knowl Discov 30(4):891\u2013927. https:\/\/doi.org\/10.1007\/s10618-015-0444-8","journal-title":"Data Mining Knowl Discov"},{"key":"1086_CR25","unstructured":"Chalapathy R, Chawla S (2019) Deep learning for anomaly detection: A survey. CoRR, arXiv: abs\/1901.03407. http:\/\/arxiv.org\/abs\/1901.03407"},{"issue":"3","key":"1086_CR26","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/1541880.1541882","volume":"41","author":"V Chandola","year":"2009","unstructured":"Chandola V, Banerjee A, Kumar V (2009) Anomaly detection. ACM Comput Surv 41(3):1\u201358. https:\/\/doi.org\/10.1145\/1541880.1541882","journal-title":"ACM Comput Surv"},{"key":"1086_CR27","doi-asserted-by":"publisher","unstructured":"Chawla S, Gionis A (2013) k-means\u2212\u2212: A unified approach to clustering and outlier detection. In Proceedings of the 13th SIAM International Conference on Data Mining (SDM), pages 189\u2013197. https:\/\/doi.org\/10.1137\/1.9781611972832.21","DOI":"10.1137\/1.9781611972832.21"},{"key":"1086_CR28","doi-asserted-by":"publisher","first-page":"209","DOI":"10.1016\/j.procs.2015.04.058","volume":"50","author":"A Christy","year":"2015","unstructured":"Christy A, Gandhi GM, Vaithyasubramanian S (2015) Cluster based outlier detection algorithm for healthcare data. Proc Comput Sci 50:209\u2013215","journal-title":"Proc Comput Sci"},{"key":"1086_CR29","doi-asserted-by":"publisher","unstructured":"Cordeiro RLF, Jr CT, Traina AJM, L\u00f3pez-Hern\u00e1ndez JC, Kang U, Faloutsos C (2011) Clustering very large multi-dimensional datasets with mapreduce. In C.\u00a0Apt\u00e9, J.\u00a0Ghosh, and P.\u00a0Smyth, editors, Proceedings of the 17th ACM SIGKDD International Conference on Knowledge Discovery and Data Mining, San Diego, CA, USA, August 21-24, 2011, pages 690\u2013698. ACM. https:\/\/doi.org\/10.1145\/2020408.2020516","DOI":"10.1145\/2020408.2020516"},{"issue":"2","key":"1086_CR30","doi-asserted-by":"publisher","first-page":"387","DOI":"10.1109\/TKDE.2011.176","volume":"25","author":"R. L. F. Cordeiro","year":"2013","unstructured":"Cordeiro R. L. F., Traina A. J. M., Faloutsos C., Traina C. (2013) Halite: fast and scalable multiresolution local-correlation clustering. IEEE Trans Knowl Data Eng 25(2):387\u2013401. https:\/\/doi.org\/10.1109\/TKDE.2011.176","journal-title":"IEEE Trans Knowl Data Eng"},{"issue":"1","key":"1086_CR31","doi-asserted-by":"publisher","first-page":"462","DOI":"10.1016\/j.jspi.2010.06.024","volume":"141","author":"P Coretto","year":"2011","unstructured":"Coretto P, Hennig C (2011) Maximum likelihood estimation of heterogeneous mixtures of gaussian and uniform distributions. J Stat Plan Inf 141(1):462\u2013473","journal-title":"J Stat Plan Inf"},{"issue":"2","key":"1086_CR32","doi-asserted-by":"publisher","first-page":"553","DOI":"10.1214\/aos\/1031833664","volume":"25","author":"JA Cuesta-Albertos","year":"1997","unstructured":"Cuesta-Albertos JA, Gordaliza A, Matr\u00e1n C (1997) Trimmed $$k$$-means: an attempt to robustify quantizers. Ann Stat 25(2):553\u2013576","journal-title":"Ann Stat"},{"key":"1086_CR33","doi-asserted-by":"crossref","unstructured":"Dang X, Assent I, Ng RT, Zimek A, Schubert E (2014) Discriminative features for identifying and interpreting outliers. In: ICDE, pages 88\u201399","DOI":"10.1109\/ICDE.2014.6816642"},{"key":"1086_CR34","doi-asserted-by":"crossref","unstructured":"Davis J, Goadrich M (2006) The relationship between precision-recall and roc curves. In: Proceedings of the 23rd international conference on Machine learning, pages 233\u2013240","DOI":"10.1145\/1143844.1143874"},{"issue":"1","key":"1086_CR35","doi-asserted-by":"publisher","first-page":"131","DOI":"10.1145\/2522968.2522981","volume":"46","author":"JA Silva","year":"2013","unstructured":"Silva JA, Faria ER, Barros RC, Hruschka ER, Carvalho ACD, Gama J (2013) Data stream clustering: a survey. ACM Comput Surv 46(1):131\u20131331. https:\/\/doi.org\/10.1145\/2522968.2522981","journal-title":"ACM Comput Surv"},{"issue":"1","key":"1086_CR36","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1111\/j.2517-6161.1977.tb01600.x","volume":"39","author":"AP Dempster","year":"1977","unstructured":"Dempster AP, Laird NM, Rubin DB (1977) Maximum likelihood from incomplete data via the EM algorithm. J Roy Stat Soc: Ser B (Methodol) 39(1):1\u201322","journal-title":"J Roy Stat Soc: Ser B (Methodol)"},{"key":"1086_CR37","first-page":"1","volume":"7","author":"J Dem\u0161ar","year":"2006","unstructured":"Dem\u0161ar J (2006) Statistical comparisons of classifiers over multiple data sets. J Mach Learn Res 7:1\u201330","journal-title":"J Mach Learn Res"},{"key":"1086_CR38","doi-asserted-by":"publisher","unstructured":"Djenouri Y, Zimek A (2018) Outlier detection in urban traffic data. In Proceedings of the 8th International Conference on Web Intelligence, Mining and Semantics, WIMS, pages 3:1\u20133:12. https:\/\/doi.org\/10.1145\/3227609.3227692","DOI":"10.1145\/3227609.3227692"},{"key":"1086_CR39","doi-asserted-by":"publisher","first-page":"406","DOI":"10.1016\/j.patcog.2017.09.037","volume":"74","author":"R Domingues","year":"2018","unstructured":"Domingues R, Filippone M, Michiardi P, Zouaoui J (2018) A comparative evaluation of outlier detection algorithms: experiments and analyses. Pattern Recognit 74:406\u2013421. https:\/\/doi.org\/10.1016\/j.patcog.2017.09.037","journal-title":"Pattern Recognit"},{"key":"1086_CR40","unstructured":"Elkan C (2003) Using the triangle inequality to accelerate k-means. In Proceedings of the Twentieth International Conference on International Conference on Machine Learning, ICML\u201903, page 147-153. AAAI Press. ISBN 1577351894"},{"key":"1086_CR41","unstructured":"Eskin E (2000) Anomaly detection over noisy data using learned probability distributions"},{"key":"1086_CR42","unstructured":"Ester M, Kriegel H-P, Sander J, Xu X (1996) A Density-Based Algorithm for Discovering Clusters in Large Spatial Databases with Noise. In International Conference on Knowledge Discovery and Data Mining, pages 226\u2013231, USA, Oregon, Portland"},{"key":"1086_CR43","unstructured":"Goldstein M (2014) Anomaly Detection in Large Datasets. PhD thesis, University of Kaiserslautern, Germany"},{"key":"1086_CR44","doi-asserted-by":"publisher","unstructured":"Goldstein M, Uchida S (2016a) A Comparative Study on Outlier Removal from a Large-scale Dataset using Unsupervised Anomaly Detection. In Proceedings of the 5th International Conference on Pattern Recognition Applications and Methods, ICPRAM, pages 263\u2013269. https:\/\/doi.org\/10.5220\/0005701302630269","DOI":"10.5220\/0005701302630269"},{"issue":"4","key":"1086_CR45","doi-asserted-by":"publisher","DOI":"10.1371\/journal.pone.0152173","volume":"11","author":"M Goldstein","year":"2016","unstructured":"Goldstein M, Uchida S (2016) A comparative evaluation of unsupervised anomaly detection algorithms for multivariate data. PLoS ONE 11(4):e0152173. https:\/\/doi.org\/10.1371\/journal.pone.0152173","journal-title":"PLoS ONE"},{"key":"1086_CR46","first-page":"235","volume":"46","author":"N G\u00f6rnitz","year":"2013","unstructured":"G\u00f6rnitz N, Kloft M, Rieck K, Brefeld U (2013) Toward supervised anomaly detection. J Artif Int Res 46:235\u2013262","journal-title":"J Artif Int Res"},{"issue":"1","key":"1086_CR47","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1080\/00401706.1969.10490657","volume":"11","author":"FE Grubbs","year":"1969","unstructured":"Grubbs FE (1969) Procedures for detecting outlying observations in samples. Technometrics 11(1):1\u201321","journal-title":"Technometrics"},{"issue":"5","key":"1086_CR48","doi-asserted-by":"publisher","first-page":"345","DOI":"10.1016\/S0306-4379(00)00022-3","volume":"25","author":"Sudipto Guha","year":"2000","unstructured":"Guha Sudipto, Rastogi Rajeev, Shim Kyuseok (2000) Rock: robust clustering algorithm for categorical attributes. Inf Syst 25(5):345\u2013366. https:\/\/doi.org\/10.1016\/S0306-4379(00)00022-3","journal-title":"Inf Syst"},{"key":"1086_CR49","doi-asserted-by":"crossref","unstructured":"Hamerly G (2010) Making k-means even faster. In Proceedings of the 2010 SIAM international conference on data mining, pages 130\u2013140. SIAM","DOI":"10.1137\/1.9781611972801.12"},{"key":"1086_CR50","doi-asserted-by":"publisher","unstructured":"Han J, Kamber M, Pei J (2012) 12 - outlier detection. In J.\u00a0Han, M.\u00a0Kamber, and J.\u00a0Pei, editors, Data Mining, The Morgan Kaufmann Series in Data Management Systems, pages 543\u2013584. Morgan Kaufmann, Boston, third edition edition. ISBN 978-0-12-381479-1. https:\/\/doi.org\/10.1016\/B978-0-12-381479-1.00012-5","DOI":"10.1016\/B978-0-12-381479-1.00012-5"},{"key":"1086_CR51","doi-asserted-by":"crossref","unstructured":"Han S, Hu X, HuangH, Jiang M, Zhao Y (2022) Adbench: Anomaly detection benchmark. In NeurIPS","DOI":"10.2139\/ssrn.4266498"},{"issue":"1","key":"1086_CR52","first-page":"100","volume":"28","author":"JA Hartigan","year":"1979","unstructured":"Hartigan JA, Wong MA (1979) Algorithm as 136: a k-means clustering algorithm. J Royal Stat Soc Series c (Appl Stat) 28(1):100\u2013108","journal-title":"J Royal Stat Soc Series c (Appl Stat)"},{"key":"1086_CR53","doi-asserted-by":"publisher","DOI":"10.1007\/978-94-015-3994-4","volume-title":"Identification of Outliers","author":"D. M. Hawkins","year":"1980","unstructured":"Hawkins D. M. (1980) Identification of Outliers. Springer Netherlands, Dordrecht. https:\/\/doi.org\/10.1007\/978-94-015-3994-4"},{"issue":"9\u201310","key":"1086_CR54","doi-asserted-by":"publisher","first-page":"1641","DOI":"10.1016\/S0167-8655(03)00003-5","volume":"24","author":"Z He","year":"2003","unstructured":"He Z, Xu X, Deng S (2003) Discovering cluster-based local outliers. Pattern Recogn Lett 24(9\u201310):1641\u20131650","journal-title":"Pattern Recogn Lett"},{"issue":"2","key":"1086_CR55","doi-asserted-by":"publisher","first-page":"85","DOI":"10.1023\/B:AIRE.0000045502.10941.a9","volume":"22","author":"VJ Hodge","year":"2004","unstructured":"Hodge VJ, Austin J (2004) A survey of outlier detection methodologies. Artif Intell Rev 22(2):85\u2013126. https:\/\/doi.org\/10.1023\/B:AIRE.0000045502.10941.a9","journal-title":"Artif Intell Rev"},{"issue":"1","key":"1086_CR56","doi-asserted-by":"publisher","first-page":"193","DOI":"10.1007\/BF01908075","volume":"2","author":"Lawrence Hubert","year":"1985","unstructured":"Hubert Lawrence, Arabie Phipps (1985) Comparing partitions. J Classif 2(1):193\u2013218. https:\/\/doi.org\/10.1007\/BF01908075","journal-title":"J Classif"},{"issue":"1\u20133","key":"1086_CR57","doi-asserted-by":"publisher","first-page":"59","DOI":"10.1007\/s10994-014-5473-9","volume":"101","author":"F Iglesias","year":"2015","unstructured":"Iglesias F, Zseby T (2015) Analysis of network traffic features for anomaly detection. Mach Learn 101(1\u20133):59\u201384. https:\/\/doi.org\/10.1007\/s10994-014-5473-9","journal-title":"Mach Learn"},{"issue":"9","key":"1086_CR58","doi-asserted-by":"publisher","first-page":"2096","DOI":"10.1109\/TPAMI.2019.2912970","volume":"42","author":"F Iglesias","year":"2020","unstructured":"Iglesias F, Zseby T, Zimek A (2020) Absolute cluster validity. IEEE Trans Pattern Anal Mach Intell 42(9):2096\u20132112. https:\/\/doi.org\/10.1109\/TPAMI.2019.2912970","journal-title":"IEEE Trans Pattern Anal Mach Intell"},{"key":"1086_CR59","doi-asserted-by":"crossref","unstructured":"Iglesias F, Zseby T, Hartl A, Zimek A (2023) Sdoclust: Clustering with sparse data observers. SISAP, pages 185\u2013199. Springer","DOI":"10.1007\/978-3-031-46994-7_16"},{"key":"1086_CR60","doi-asserted-by":"crossref","unstructured":"Iglesias V\u00e1zquez F, Zseby T, Zimek A (2018) Outlier detection based on low density models. In ICDM Workshops, pages 970\u2013979. IEEE","DOI":"10.1109\/ICDMW.2018.00140"},{"key":"1086_CR61","doi-asserted-by":"publisher","DOI":"10.1016\/j.eswa.2023.120994","volume":"233","author":"F Iglesias V\u00e1zquez","year":"2023","unstructured":"Iglesias V\u00e1zquez F, Hartl A, Zseby T, Zimek A (2023) Anomaly detection in streaming data: a comparison and evaluation study. Expert Syst Appl 233:120994. https:\/\/doi.org\/10.1016\/j.eswa.2023.120994","journal-title":"Expert Syst Appl"},{"issue":"2","key":"1086_CR62","doi-asserted-by":"publisher","first-page":"329","DOI":"10.1007\/s10115-015-0851-6","volume":"47","author":"PA Jaskowiak","year":"2016","unstructured":"Jaskowiak PA, Moulavi D, Furtado ACS, Campello RJGB, Zimek A, Sander J (2016) On strategies for building effective ensembles of relative clustering validity criteria. Knowl Inf Syst 47(2):329\u2013354. https:\/\/doi.org\/10.1007\/s10115-015-0851-6","journal-title":"Knowl Inf Syst"},{"issue":"3","key":"1086_CR63","doi-asserted-by":"publisher","first-page":"1219","DOI":"10.1007\/s10618-022-00829-0","volume":"36","author":"PA Jaskowiak","year":"2022","unstructured":"Jaskowiak PA, Costa IG, Campello RJGB (2022) The area under the ROC curve as a measure of clustering quality. Data Min Knowl Discov 36(3):1219\u20131245. https:\/\/doi.org\/10.1007\/s10618-022-00829-0","journal-title":"Data Min Knowl Discov"},{"issue":"6\u20137","key":"1086_CR64","doi-asserted-by":"publisher","first-page":"691","DOI":"10.1016\/S0167-8655(00)00131-8","volume":"22","author":"M-F Jiang","year":"2001","unstructured":"Jiang M-F, Tseng S-S, Su C-M (2001) Two-phase clustering process for outliers detection. Pattern Recogn Lett 22(6\u20137):691\u2013700","journal-title":"Pattern Recogn Lett"},{"key":"1086_CR65","doi-asserted-by":"crossref","unstructured":"Jin W, Tung AK, Han J, Wang W (2006) Ranking outliers using symmetric neighborhood relationship. In: Pacific-Asia conference on knowledge discovery and data mining, pages 577\u2013593. Springer","DOI":"10.1007\/11731139_68"},{"issue":"3","key":"1086_CR66","doi-asserted-by":"publisher","first-page":"12","DOI":"10.4018\/IJSSCI.2021070102","volume":"13","author":"SM Kaddour","year":"2021","unstructured":"Kaddour SM, Lehsaini M (2021) Electricity consumption data analysis using various outlier detection methods. Int J Softw Sci Comput Int(IJSSCI) 13(3):12\u201327","journal-title":"Int J Softw Sci Comput Int(IJSSCI)"},{"key":"1086_CR67","doi-asserted-by":"publisher","unstructured":"Keogh E, Lin J, Fu A (2005) Hot sax: Efficiently finding the most unusual time series subsequence. In Proceedings of the Fifth IEEE International Conference on Data Mining, ICDM \u201905, page 226-233, USA. IEEE Computer Society. ISBN 0769522785. https:\/\/doi.org\/10.1109\/ICDM.2005.79","DOI":"10.1109\/ICDM.2005.79"},{"key":"1086_CR68","unstructured":"Knorr EM, Ng RT (1998) Algorithms for mining distance-based outliers in large datasets. In: VLDB, pages 392\u2013403"},{"key":"1086_CR69","unstructured":"Knorr EM, Ng RT (1999) Finding intensional knowledge of distance-based outliers. In: VLDB, pages 211\u2013222"},{"key":"1086_CR70","doi-asserted-by":"publisher","unstructured":"Kriegel H, Schubert M, Zimek A (2008) Angle-based outlier detection in high-dimensional data. In: KDD, pages 444\u2013452. ACM. https:\/\/doi.org\/10.1145\/1401890.1401946","DOI":"10.1145\/1401890.1401946"},{"key":"1086_CR71","doi-asserted-by":"publisher","first-page":"831","DOI":"10.1007\/978-3-642-01307-2_86","volume-title":"Advances in Knowledge Discovery and Data Mining","author":"Hans-Peter Kriegel","year":"2009","unstructured":"Kriegel Hans-Peter, Kr\u00f6ger Peer, Schubert Erich, Zimek Arthur (2009) Outlier Detection in Axis-Parallel Subspaces of High Dimensional Data. In: Theeramunkong Thanaruk, Kijsirikul Boonserm, Cercone Nick, Ho Tu-Bao (eds) Advances in Knowledge Discovery and Data Mining. Springer Berlin Heidelberg, Berlin, Heidelberg, pp 831\u2013838. https:\/\/doi.org\/10.1007\/978-3-642-01307-2_86"},{"issue":"1","key":"1086_CR72","doi-asserted-by":"publisher","first-page":"1:1","DOI":"10.1145\/1497577.1497578","volume":"3","author":"H Kriegel","year":"2009","unstructured":"Kriegel H, Kr\u00f6ger P, Zimek A (2009) Clustering high-dimensional data: a survey on subspace clustering, pattern-based clustering, and correlation clustering. ACM Trans Knowl Discov Data 3(1):1:1-1:58","journal-title":"ACM Trans Knowl Discov Data"},{"key":"1086_CR73","doi-asserted-by":"crossref","unstructured":"Kriegel H, Kr\u00f6ger P, Schubert E, Zimek A (2012) Outlier detection in arbitrarily oriented subspaces. In ICDM, pages 379\u2013388","DOI":"10.1109\/ICDM.2012.21"},{"issue":"2","key":"1086_CR74","doi-asserted-by":"publisher","first-page":"341","DOI":"10.1007\/s10115-016-1004-2","volume":"52","author":"H Kriegel","year":"2017","unstructured":"Kriegel H, Schubert E, Zimek A (2017) The (black) art of runtime evaluation: are we comparing algorithms or implementations? Knowl Inf Syst 52(2):341\u2013378. https:\/\/doi.org\/10.1007\/s10115-016-1004-2","journal-title":"Knowl Inf Syst"},{"key":"1086_CR75","doi-asserted-by":"publisher","first-page":"22","DOI":"10.1016\/j.is.2014.03.001","volume":"44","author":"HD Kuna","year":"2014","unstructured":"Kuna HD, Garc\u00eda-Martinez R, Villatoro FR (2014) Outlier detection in audit logs for application systems. Inf Syst 44:22\u201333","journal-title":"Inf Syst"},{"key":"1086_CR76","unstructured":"Lane D (2003) Online statistics education: A multimedia course of study. Association for the Advancement of Computing in Education (AACE)"},{"key":"1086_CR77","doi-asserted-by":"publisher","unstructured":"Lenssen L, Schubert E (2022) Clustering by direct optimization of the medoid silhouette. In: Proceedings of the 15th International Conference on Similarity Search and Applications, SISAP, pages 190\u2013204. https:\/\/doi.org\/10.1007\/978-3-031-17849-8_15","DOI":"10.1007\/978-3-031-17849-8_15"},{"issue":"1","key":"1086_CR78","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/3609333","volume":"18","author":"Zhong Li","year":"2024","unstructured":"Li Zhong, Zhu Yuxuan, Van Leeuwen Matthijs (2024) A survey on explainable anomaly detection. ACM Trans Knowl Discov Data 18(1):1\u201354. https:\/\/doi.org\/10.1145\/3609333","journal-title":"ACM Trans Knowl Discov Data"},{"key":"1086_CR79","volume-title":"Mixture Models: Theory, Geometry and Applications","author":"Bruce G. Lindsay","year":"1995","unstructured":"Lindsay Bruce G. (1995) Mixture Models: Theory, Geometry and Applications. Institute of Mathematical Statistics and American Statistical Association, Haywood CA and Alexandria VA"},{"key":"1086_CR80","doi-asserted-by":"crossref","unstructured":"Liu FT, Ting KM, Zhou Z-H (2008) Isolation forest. In: 2008 eighth ieee international conference on data mining, pages 413\u2013422. IEEE","DOI":"10.1109\/ICDM.2008.17"},{"issue":"1","key":"1086_CR81","doi-asserted-by":"publisher","first-page":"19","DOI":"10.1145\/3606274.3606277","volume":"25","author":"MQ Ma","year":"2023","unstructured":"Ma MQ, Zhao Y, Zhang X, Akoglu L (2023) The need for unsupervised outlier model selection: a review and evaluation of internal evaluation strategies. SIGKDD Explor 25(1):19\u201335","journal-title":"SIGKDD Explor"},{"key":"1086_CR82","unstructured":"MacQueen J (1967) Some methods for classification and analysis of multivariate observation. In: Proceedings of the 5th Berkley Symposium on Mathematical Statistics and Probability, pages 281\u2013297"},{"issue":"4","key":"1086_CR83","doi-asserted-by":"publisher","first-page":"167","DOI":"10.1002\/sam.11380","volume":"11","author":"S Mansalis","year":"2018","unstructured":"Mansalis S, Ntoutsi E, Pelekis N, Theodoridis Y (2018) An evaluation of data stream clustering algorithms. Stat Anal Data Min 11(4):167\u2013187. https:\/\/doi.org\/10.1002\/sam.11380","journal-title":"Stat Anal Data Min"},{"issue":"4","key":"1086_CR84","doi-asserted-by":"publisher","first-page":"47:1","DOI":"10.1145\/3394053","volume":"14","author":"HO Marques","year":"2020","unstructured":"Marques HO, Campello RJGB, Sander J, Zimek A (2020) Internal evaluation of unsupervised outlier detection. ACM Trans Knowl Discov Data 14(4):47:1-47:42. https:\/\/doi.org\/10.1145\/3394053","journal-title":"ACM Trans Knowl Discov Data"},{"key":"1086_CR85","doi-asserted-by":"publisher","unstructured":"Marques HO, Zimek A, Campello RJ GB, Sander J (2022) Similarity-based unsupervised evaluation of outlier detection. In: SISAP, pages 234\u2013248. https:\/\/doi.org\/10.1007\/978-3-031-17849-8_19","DOI":"10.1007\/978-3-031-17849-8_19"},{"key":"1086_CR86","doi-asserted-by":"publisher","first-page":"1473","DOI":"10.1007\/s10618-023-00931-x","volume":"37","author":"HO Marques","year":"2023","unstructured":"Marques HO, Swersky L, Sander J, Campello RJGB, Zimek A (2023) On the evaluation of outlier detection and one-class classification: a comparative study of algorithms, model selection, and ensembles. Data Min Knowl Discov 37:1473\u20131517. https:\/\/doi.org\/10.1007\/s10618-023-00931-x","journal-title":"Data Min Knowl Discov"},{"key":"1086_CR87","doi-asserted-by":"publisher","unstructured":"Moise G, Sander J, Ester M (2006) P3c: A robust projected clustering algorithm. In Sixth International Conference on Data Mining (ICDM\u201906), pages 414\u2013425. https:\/\/doi.org\/10.1109\/ICDM.2006.123","DOI":"10.1109\/ICDM.2006.123"},{"issue":"01","key":"1086_CR88","doi-asserted-by":"publisher","first-page":"19","DOI":"10.1142\/S0218213008003753","volume":"17","author":"H Moonesinghe","year":"2008","unstructured":"Moonesinghe H, Tan P-N (2008) Outrank: a graph-based outlier detection framework using random walk. Int J Artif Intell Tools 17(01):19\u201336","journal-title":"Int J Artif Intell Tools"},{"key":"1086_CR89","doi-asserted-by":"publisher","unstructured":"Moulavi D, Jaskowiak PA, Campello RJ, Zimek A, Sander J (2014) Density-based clustering validation. SIAM International Conference on Data Mining 2014, SDM 2014, 2(i), 839\u2013847. https:\/\/doi.org\/10.1137\/1.9781611973440.96","DOI":"10.1137\/1.9781611973440.96"},{"key":"1086_CR90","doi-asserted-by":"publisher","unstructured":"M\u00fcller E, Assent I, Steinhausen U, Seidl T (2008) OutRank: Ranking outliers in high dimensional data. In: Proceedings - International Conference on Data Engineering, pages 600\u2013603. ISSN 10844627. https:\/\/doi.org\/10.1109\/ICDEW.2008.4498387","DOI":"10.1109\/ICDEW.2008.4498387"},{"key":"1086_CR91","doi-asserted-by":"publisher","unstructured":"M\u00fcller E, Schiffer M, Seidl T (2010) Adaptive outlierness for subspace outlier ranking. In: Proceedings of the 19th ACM Conference on Information and Knowledge Management (CIKM), Toronto, ON, Canada, pages 1629\u20131632. https:\/\/doi.org\/10.1145\/1871437.1871690","DOI":"10.1145\/1871437.1871690"},{"key":"1086_CR92","doi-asserted-by":"publisher","unstructured":"M\u00fcller E, Schiffer M, Seidl T (2011) Statistical selection of relevant subspace projections for outlier ranking. In: Proceedings of the 27th International Conference on Data Engineering (ICDE), Hannover, Germany, pages 434\u2013445. https:\/\/doi.org\/10.1109\/ICDE.2011.5767916","DOI":"10.1109\/ICDE.2011.5767916"},{"issue":"2","key":"1086_CR93","doi-asserted-by":"publisher","first-page":"259","DOI":"10.1007\/s10618-012-0290-x","volume":"27","author":"MC Naldi","year":"2013","unstructured":"Naldi MC, de Carvalho ACPLF, Campello RJGB (2013) Cluster ensemble selection based on relative validity indexes. Data Min Knowl Discov 27(2):259\u2013289. https:\/\/doi.org\/10.1007\/s10618-012-0290-x","journal-title":"Data Min Knowl Discov"},{"key":"1086_CR94","unstructured":"Ng R, Han J (1994) Efficient and effective clustering method for spatial data mining. In: Proc. of the 20th VLDB Conference, pages 144\u2013155"},{"issue":"3","key":"1086_CR95","doi-asserted-by":"publisher","first-page":"559","DOI":"10.1016\/j.dss.2010.08.006","volume":"50","author":"EW Ngai","year":"2011","unstructured":"Ngai EW, Hu Y, Wong YH, Chen Y, Sun X (2011) The application of data mining techniques in financial fraud detection: a classification framework and an academic review of literature. Decis Support Syst 50(3):559\u2013569","journal-title":"Decis Support Syst"},{"key":"1086_CR96","doi-asserted-by":"publisher","unstructured":"Okkels CB, Aum\u00fcller M, Zimek A (2024) On the design of scalable outlier detection methods using approximate nearest neighbor graphs. In: SISAP, pages 170\u2013184. https:\/\/doi.org\/10.1007\/978-3-031-75823-2_14","DOI":"10.1007\/978-3-031-75823-2_14"},{"issue":"1\u20132","key":"1086_CR97","first-page":"1469","volume":"3","author":"GH Orair","year":"2010","unstructured":"Orair GH, Teixeira CHC, Meira W, Wang Y, Parthasarathy S (2010) Distance-based outlier detection: consolidation and renewed bearing. VLDB 3(1\u20132):1469\u20131480","journal-title":"VLDB"},{"issue":"2","key":"1086_CR98","doi-asserted-by":"publisher","first-page":"38:1","DOI":"10.1145\/3439950","volume":"54","author":"G Pang","year":"2022","unstructured":"Pang G, Shen C, Cao L, van den Hengel A (2022) Deep learning for anomaly detection: a review. ACM Comput Surv 54(2):38:1-38:38. https:\/\/doi.org\/10.1145\/3439950","journal-title":"ACM Comput Surv"},{"key":"1086_CR99","doi-asserted-by":"publisher","unstructured":"Papadimitriou S, Kitagawa H, Gibbons P, Faloutsos C (2003) LOCI: fast outlier detection using the local correlation integral. In: Proceedings 19th International Conference on Data Engineering, number\u00a01, pages 315\u2013326. https:\/\/doi.org\/10.1109\/ICDE.2003.1260802","DOI":"10.1109\/ICDE.2003.1260802"},{"key":"1086_CR100","doi-asserted-by":"crossref","unstructured":"Ramaswamy S, Rastogi R, Shim K (2000) Efficient Algorithms for Mining Outliers from Large Data Sets. In Proceedings of the 2000 ACM SIGMOD international conference on Management of data, pages 427\u2013438. ACM SIGMOD","DOI":"10.1145\/342009.335437"},{"issue":"4","key":"1086_CR101","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/2890508","volume":"10","author":"S Rayana","year":"2016","unstructured":"Rayana S, Akoglu L (2016) Less is more: building selective anomaly ensembles. ACM Trans Knowl Discov Data (tkdd) 10(4):1\u201333","journal-title":"ACM Trans Knowl Discov Data (tkdd)"},{"key":"1086_CR102","doi-asserted-by":"publisher","first-page":"53","DOI":"10.1016\/0377-0427(87)90125-7","volume":"20","author":"PJ Rousseeuw","year":"1987","unstructured":"Rousseeuw PJ (1987) Silhouettes: a graphical aid to the interpretation and validation of cluster analysis. J Comput Appl Math 20:53\u201365","journal-title":"J Comput Appl Math"},{"issue":"1","key":"1086_CR103","doi-asserted-by":"publisher","first-page":"33","DOI":"10.1145\/2594473.2594479","volume":"15","author":"MS Sadik","year":"2013","unstructured":"Sadik MS, Gruenwald L (2013) Research issues in outlier detection for data streams. SIGKDD Explor 15(1):33\u201340. https:\/\/doi.org\/10.1145\/2594473.2594479","journal-title":"SIGKDD Explor"},{"key":"1086_CR104","doi-asserted-by":"crossref","unstructured":"Schlegl T, Seeb\u00f6ck P, Waldstein SM, Schmidt-Erfurth U, Langs G (2017) Unsupervised anomaly detection with generative adversarial networks to guide marker discovery. In: Proceedings of the 25th International Conference on Information Processing in Medical Imaging, IPMI, pages 146\u2013157. Springer","DOI":"10.1007\/978-3-319-59050-9_12"},{"key":"1086_CR105","doi-asserted-by":"publisher","unstructured":"Schubert E (2022) Automatic indexing for similarity search in ELKI. In Proceedings of the 15th International Conference on Similarity Search and Applications, SISAP, pages 205\u2013213. https:\/\/doi.org\/10.1007\/978-3-031-17849-8_16","DOI":"10.1007\/978-3-031-17849-8_16"},{"key":"1086_CR106","doi-asserted-by":"publisher","unstructured":"Schubert E, Weiler M, Zimek A (2015a) Outlier detection and trend detection: Two sides of the same coin. In IEEE International Conference on Data Mining Workshop, ICDMW, pages 40\u201346.https:\/\/doi.org\/10.1109\/ICDMW.2015.79","DOI":"10.1109\/ICDMW.2015.79"},{"key":"1086_CR107","first-page":"19","volume":"2","author":"E Schubert","year":"2015","unstructured":"Schubert E, Zimek A, Kriegel H (2015) Fast and scalable outlier detection with approximate nearest neighbor ensembles. In DASFAA 2:19\u201336","journal-title":"In DASFAA"},{"key":"1086_CR108","doi-asserted-by":"publisher","unstructured":"Silva AED, Sanches LL, Fraideinberze AC, Cordeiro RLF (2016) $$\\mathit{Halite}_{\\text{ds}}$$: Fast and scalable subspace clustering for multidimensional data streams. In S.\u00a0C. Venkatasubramanian and W.\u00a0M. Jr., editors, Proceedings of the 2016 SIAM International Conference on Data Mining, Miami, Florida, USA, May 5-7, 2016, pages 351\u2013359. SIAM. https:\/\/doi.org\/10.1137\/1.9781611974348.40","DOI":"10.1137\/1.9781611974348.40"},{"key":"1086_CR109","doi-asserted-by":"crossref","unstructured":"Su W, Yuan Y, Zhu M (2015) A relationship between the average precision and the area under the roc curve. In: Proceedings of the 2015 international conference on the theory of information retrieval, pages 349\u2013352","DOI":"10.1145\/2808194.2809481"},{"key":"1086_CR110","doi-asserted-by":"crossref","unstructured":"Tang J, Chen Z, Fu AW-C, Cheung DW (2002) Enhancing effectiveness of outlier detections for low density patterns. In: Proceedings of the 6th Pacific-Asia Conference on Advances in Knowledge Discovery and Data Mining, PAKDD, pages 535\u2013548","DOI":"10.1007\/3-540-47887-6_53"},{"issue":"8","key":"1086_CR111","doi-asserted-by":"publisher","first-page":"161","DOI":"10.14257\/ijca.2015.8.8.17","volume":"8","author":"X-M Tang","year":"2015","unstructured":"Tang X-M, Yuan R-X, Chen J (2015) Outlier detection in energy disaggregation using subspace learning and gaussian mixture model. Int J Control Automat 8(8):161\u2013170","journal-title":"Int J Control Automat"},{"issue":"8","key":"1086_CR112","doi-asserted-by":"publisher","first-page":"575","DOI":"10.1080\/0094965031000136012","volume":"73","author":"MJ Van der Laan","year":"2003","unstructured":"Van der Laan MJ, Pollard KS, Bryan J (2003) A new partitioning around medoids algorithm. J Stat Comput Simul 73(8):575\u2013584. https:\/\/doi.org\/10.1080\/0094965031000136012","journal-title":"J Stat Comput Simul"},{"issue":"4","key":"1086_CR113","doi-asserted-by":"publisher","first-page":"209","DOI":"10.1002\/sam.10080","volume":"3","author":"Lucas Vendramin","year":"2010","unstructured":"Vendramin Lucas, Campello Ricardo J. G. B., Hruschka Eduardo R. (2010) Relative clustering validity criteria: a comparative overview. Stat Anal Data Mining ASA Data Sci J 3(4):209\u2013235. https:\/\/doi.org\/10.1002\/sam.10080","journal-title":"Stat Anal Data Mining ASA Data Sci J"},{"key":"1086_CR114","doi-asserted-by":"publisher","first-page":"75531","DOI":"10.1109\/ACCESS.2018.2883681","volume":"6","author":"C Wang","year":"2018","unstructured":"Wang C, Gao H, Liu Z, Fu Y (2018) A new outlier detection model using random walk on local information graph. IEEE Access 6:75531\u201375544","journal-title":"IEEE Access"},{"key":"1086_CR115","doi-asserted-by":"crossref","unstructured":"Wang Y, Parthasarathy S, Tatikonda S (2011) Locality sensitive outlier detection: a ranking driven approach. In: ICDE, pages 410\u2013421","DOI":"10.1109\/ICDE.2011.5767852"},{"key":"1086_CR116","doi-asserted-by":"crossref","unstructured":"Yamanishi K, Takeuchi J-I, Williams G, Milne P (2000) On-line unsupervised outlier detection using finite mixtures with discounting learning algorithms. In: Proceedings of the sixth ACM SIGKDD international conference on Knowledge discovery and data mining, pages 320\u2013324","DOI":"10.1145\/347090.347160"},{"key":"1086_CR117","doi-asserted-by":"crossref","unstructured":"Yang X, Latecki LJ, Pokrajac D (2009) Outlier detection with globally optimal exemplar-based gmm. In: Proceedings of the 2009 SIAM international conference on data mining, pages 145\u2013154. SIAM","DOI":"10.1137\/1.9781611972795.13"},{"issue":"4","key":"1086_CR118","doi-asserted-by":"publisher","first-page":"506","DOI":"10.1109\/TBDATA.2017.2672672","volume":"5","author":"W Yu","year":"2017","unstructured":"Yu W, Li J, Bhuiyan MZA, Zhang R, Huai J (2017) Ring: real-time emerging anomaly monitoring system over text streams. IEEE Trans Big Data 5(4):506\u2013519","journal-title":"IEEE Trans Big Data"},{"issue":"2","key":"1086_CR119","doi-asserted-by":"publisher","first-page":"103","DOI":"10.1145\/235968.233324","volume":"25","author":"T Zhang","year":"1996","unstructured":"Zhang T, Ramakrishnan R, Livny M (1996) Birch: an efficient data clustering method for very large databases. ACM SIGMOD Rec 25(2):103\u2013114","journal-title":"ACM SIGMOD Rec"},{"issue":"6","key":"1086_CR120","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1002\/widm.1280","volume":"8","author":"A Zimek","year":"2018","unstructured":"Zimek A, Filzmoser P (2018) There and back again: outlier detection between statistical reasoning and data mining algorithms. WIREs Data Mining Knowl Discov 8(6):1\u201326. https:\/\/doi.org\/10.1002\/widm.1280","journal-title":"WIREs Data Mining Knowl Discov"},{"key":"1086_CR121","doi-asserted-by":"publisher","unstructured":"Zimek A, Schubert E (2018) Outlier detection. In L.\u00a0Liu and M.\u00a0T. \u00d6zsu, editors, Encyclopedia of Database Systems, Second Edition. Springer. https:\/\/doi.org\/10.1007\/978-1-4614-8265-9_80719","DOI":"10.1007\/978-1-4614-8265-9_80719"},{"issue":"5","key":"1086_CR122","doi-asserted-by":"publisher","first-page":"363","DOI":"10.1002\/sam.11161","volume":"5","author":"A Zimek","year":"2012","unstructured":"Zimek A, Schubert E, Kriegel H-P (2012) A survey on unsupervised outlier detection in high-dimensional numerical data. Stat Anal Data Mining 5(5):363\u2013387. https:\/\/doi.org\/10.1002\/sam.11161","journal-title":"Stat Anal Data Mining"},{"issue":"1","key":"1086_CR123","doi-asserted-by":"publisher","first-page":"11","DOI":"10.1145\/2594473.2594476","volume":"15","author":"A Zimek","year":"2014","unstructured":"Zimek A, Campello RJ, Sander J (2014) Ensembles for unsupervised outlier detection. ACM SIGKDD Explor Newslett 15(1):11\u201322. https:\/\/doi.org\/10.1145\/2594473.2594476","journal-title":"ACM SIGKDD Explor Newslett"},{"issue":"2","key":"1086_CR124","doi-asserted-by":"publisher","first-page":"1201","DOI":"10.1007\/s10462-020-09874-x","volume":"54","author":"Alaettin Zubaro\u011flu","year":"2021","unstructured":"Zubaro\u011flu Alaettin, Atalay Volkan (2021) Data stream clustering: a review. Artif Int Rev 54(2):1201\u20131236. https:\/\/doi.org\/10.1007\/s10462-020-09874-x","journal-title":"Artif Int Rev"}],"container-title":["Data Mining and Knowledge Discovery"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s10618-024-01086-z.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s10618-024-01086-z\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s10618-024-01086-z.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,3,10]],"date-time":"2025-03-10T02:48:31Z","timestamp":1741574911000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s10618-024-01086-z"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,2,3]]},"references-count":124,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2025,3]]}},"alternative-id":["1086"],"URL":"https:\/\/doi.org\/10.1007\/s10618-024-01086-z","relation":{},"ISSN":["1384-5810","1573-756X"],"issn-type":[{"value":"1384-5810","type":"print"},{"value":"1573-756X","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,2,3]]},"assertion":[{"value":"15 December 2024","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"28 December 2024","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"3 February 2025","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"Arthur Zimek is serving as action editor for DAMI.","order":2,"name":"Ethics","group":{"name":"EthicsHeading","label":"Conflict of interest"}}],"article-number":"13"}}