{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,2,24]],"date-time":"2026-02-24T16:19:03Z","timestamp":1771949943225,"version":"3.50.1"},"reference-count":58,"publisher":"Association for Computing Machinery (ACM)","issue":"2","license":[{"start":{"date-parts":[[2006,4,1]],"date-time":"2006-04-01T00:00:00Z","timestamp":1143849600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Inf. Syst."],"published-print":{"date-parts":[[2006,4]]},"abstract":"<jats:p>With continued advances in communication network technology and sensing technology, there is astounding growth in the amount of data produced and made available through cyberspace. Efficient and high-quality clustering of large datasets continues to be one of the most important problems in large-scale data analysis. A commonly used methodology for cluster analysis on large datasets is the three-phase framework of sampling\/summarization, iterative cluster analysis, and disk-labeling. There are three known problems with this framework which demand effective solutions. The first problem is how to effectively define and validate irregularly shaped clusters, especially in large datasets. Automated algorithms and statistical methods are typically not effective in handling these particular clusters. The second problem is how to effectively label the entire data on disk (disk-labeling) without introducing additional errors, including the solutions for dealing with outliers, irregular clusters, and cluster boundary extension. The third obstacle is the lack of research about issues related to effectively integrating the three phases. In this article, we describe iVIBRATE---an interactive visualization-based three-phase framework for clustering large datasets. The two main components of iVIBRATE are its VISTA visual cluster-rendering subsystem which invites human interplay into the large-scale iterative clustering process through interactive visualization, and its adaptive ClusterMap labeling subsystem which offers visualization-guided disk-labeling solutions that are effective in dealing with outliers, irregular clusters, and cluster boundary extension. Another important contribution of iVIBRATE development is the identification of the special issues presented in integrating the two components and the sampling approach into a coherent framework, as well as the solutions for improving the reliability of the framework and for minimizing the amount of errors generated within the cluster analysis process. We study the effectiveness of the iVIBRATE framework through a walkthrough example dataset of a million records and we experimentally evaluate the iVIBRATE approach using both real-life and synthetic datasets. Our results show that iVIBRATE can efficiently involve the user in the clustering process and generate high-quality clustering results for large datasets.<\/jats:p>","DOI":"10.1145\/1148020.1148024","type":"journal-article","created":{"date-parts":[[2006,10,18]],"date-time":"2006-10-18T18:11:32Z","timestamp":1161195092000},"page":"245-294","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":39,"title":["iVIBRATE"],"prefix":"10.1145","volume":"24","author":[{"given":"Keke","family":"Chen","sequence":"first","affiliation":[{"name":"Georgia Institute of Technology, Atlanta, GA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Ling","family":"Liu","sequence":"additional","affiliation":[{"name":"Georgia Institute of Technology, Atlanta, GA"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2006,4]]},"reference":[{"key":"e_1_2_1_1_1","volume-title":"Proceedings of the ACM SIGMOD Conference, 49--60","author":"Ankerst M.","unstructured":"Ankerst , M. , Breunig , M. M. , Kriegel , H.-P. , and Sander , J . 1999. OPTICS: Ordering points to identify the clustering structure . In Proceedings of the ACM SIGMOD Conference, 49--60 .]] 10.1145\/304182.304187 Ankerst, M., Breunig, M. M., Kriegel, H.-P., and Sander, J. 1999. OPTICS: Ordering points to identify the clustering structure. In Proceedings of the ACM SIGMOD Conference, 49--60.]] 10.1145\/304182.304187"},{"key":"e_1_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1137\/0906011"},{"key":"e_1_2_1_3_1","volume-title":"Modern Information Retrieval","author":"Baeza-Yates R.","unstructured":"Baeza-Yates , R. and Ribeiro-Neto , B. 1999. Modern Information Retrieval . Addison Wesley , New York .]] Baeza-Yates, R. and Ribeiro-Neto, B. 1999. Modern Information Retrieval. Addison Wesley, New York.]]"},{"key":"e_1_2_1_4_1","first-page":"3","article-title":"The New Jersey data reduction report","volume":"20","author":"Barbar\u00e1 D.","year":"1997","unstructured":"Barbar\u00e1 , D. , DuMouchel , W. , Faloutsos , C. , Haas , P. J. , Hellerstein , J. M. , Ioannidis , Y. E. , Jagadish , H. V. , Johnson , T. , Ng , R. T. , Poosala , V. , Ross , K. A. , and Sevcik , K. C. 1997 . The New Jersey data reduction report . IEEE Data Eng. Bull. 20 , 4, 3 -- 45 .]] Barbar\u00e1, D., DuMouchel, W., Faloutsos, C., Haas, P. J., Hellerstein, J. M., Ioannidis, Y. E., Jagadish, H. V., Johnson, T., Ng, R. T., Poosala, V., Ross, K. A., and Sevcik, K. C. 1997. The New Jersey data reduction report. IEEE Data Eng. Bull. 20, 4, 3--45.]]","journal-title":"IEEE Data Eng. Bull."},{"key":"e_1_2_1_5_1","volume-title":"Proceedings of the ACM SIGKDD Conference, 9--15","author":"Bradley P. S.","unstructured":"Bradley , P. S. , Fayyad , U. M. , and Reina , C . 1998. Scaling clustering algorithms to large databases . In Proceedings of the ACM SIGKDD Conference, 9--15 .]] Bradley, P. S., Fayyad, U. M., and Reina, C. 1998. Scaling clustering algorithms to large databases. In Proceedings of the ACM SIGKDD Conference, 9--15.]]"},{"key":"e_1_2_1_6_1","volume-title":"Proceedings of the ACM Conference on Information and Knowledge Management (CIKM), 285--293","author":"Chen K.","unstructured":"Chen , K. and Liu , L . 2004a. Clustermap: Labeling clusters in large datasets via visualization . In Proceedings of the ACM Conference on Information and Knowledge Management (CIKM), 285--293 .]] 10.1145\/1031171.1031233 Chen, K. and Liu, L. 2004a. Clustermap: Labeling clusters in large datasets via visualization. In Proceedings of the ACM Conference on Information and Knowledge Management (CIKM), 285--293.]] 10.1145\/1031171.1031233"},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1057\/palgrave.ivs.9500076"},{"key":"e_1_2_1_8_1","volume-title":"Proceedings of the International Conference on Scientific and Statistical Database Management (SSDBM), 253--262","author":"Chen K.","unstructured":"Chen , K. and Liu , L. 2005. The \u201cbest k\u201d for entropy-based categorical clustering . In Proceedings of the International Conference on Scientific and Statistical Database Management (SSDBM), 253--262 .]] Chen, K. and Liu, L. 2005. The \u201cbest k\u201d for entropy-based categorical clustering. In Proceedings of the International Conference on Scientific and Statistical Database Management (SSDBM), 253--262.]]"},{"key":"e_1_2_1_9_1","volume-title":"Practical Nonparametric Statistics","author":"Conover W.","unstructured":"Conover , W. 1998. Practical Nonparametric Statistics . John Wiley and Sons , New York .]] Conover, W. 1998. Practical Nonparametric Statistics. John Wiley and Sons, New York.]]"},{"key":"e_1_2_1_10_1","doi-asserted-by":"crossref","first-page":"155","DOI":"10.1080\/10618600.1995.10474674","article-title":"Grand tour and projection pursuit","volume":"23","author":"Cook D.","year":"1995","unstructured":"Cook , D. , Buja , A. , Cabrera , J. , and Hurley , C. 1995 . Grand tour and projection pursuit . J. Comput. Graphical Statistics 23 , 155 -- 172 .]] Cook, D., Buja, A., Cabrera, J., and Hurley, C. 1995. Grand tour and projection pursuit. J. Comput. Graphical Statistics 23, 155--172.]]","journal-title":"J. Comput. Graphical Statistics"},{"key":"e_1_2_1_11_1","doi-asserted-by":"crossref","unstructured":"Cox T. F. and Cox M. A. A. 2001. Multidimensional Scaling. Chapman and Hall London UK.]] Cox T. F. and Cox M. A. A. 2001. Multidimensional Scaling. Chapman and Hall London UK.]]","DOI":"10.1201\/9780367801700"},{"key":"e_1_2_1_12_1","volume-title":"Proceedings of the ACM SIGIR Conference, 318--329","author":"Cutting D. R.","unstructured":"Cutting , D. R. , Karger , D. R. , Pedersen , J. O. , and Tukey , J. W . 1992. Scater\/Gather: A cluster-based approach to browsing large document collections . In Proceedings of the ACM SIGIR Conference, 318--329 .]] 10.1145\/133160.133214 Cutting, D. R., Karger, D. R., Pedersen, J. O., and Tukey, J. W. 1992. Scater\/Gather: A cluster-based approach to browsing large document collections. In Proceedings of the ACM SIGIR Conference, 318--329.]] 10.1145\/133160.133214"},{"key":"e_1_2_1_13_1","first-page":"488","article-title":"Visualizing class structure of multidimensional data. In Proceedings of the 30th Symposium on the Interface","volume":"30","author":"Dhillon I. S.","year":"1998","unstructured":"Dhillon , I. S. , Modha , D. S. , and Spangler , W. S. 1998 . Visualizing class structure of multidimensional data. In Proceedings of the 30th Symposium on the Interface : Computing Science and Statistics 30 , 488 -- 493 .]] Dhillon, I. S., Modha, D. S., and Spangler, W. S. 1998. Visualizing class structure of multidimensional data. In Proceedings of the 30th Symposium on the Interface: Computing Science and Statistics 30, 488--493.]]","journal-title":"Computing Science and Statistics"},{"key":"e_1_2_1_14_1","article-title":"Validity studies in clustering methodologies. Pattern","author":"Dubes R. C.","year":"1979","unstructured":"Dubes , R. C. and Jain , A. K. 1979 . Validity studies in clustering methodologies. Pattern Recogn. Lett., 235--254.]] Dubes, R. C. and Jain, A. K. 1979. Validity studies in clustering methodologies. Pattern Recogn. Lett., 235--254.]]","journal-title":"Recogn. Lett., 235--254.]]"},{"key":"e_1_2_1_15_1","volume-title":"Proceedings of the 2nd International Conference on Knowledge Discovery and Data Mining, 226--231","author":"Ester M.","unstructured":"Ester , M. , Kriegel , H.-P. , Sander , J. , and Xu , X . 1996. A density-based algorithm for discovering clusters in large spatial databases with noise . In Proceedings of the 2nd International Conference on Knowledge Discovery and Data Mining, 226--231 .]] Ester, M., Kriegel, H.-P., Sander, J., and Xu, X. 1996. A density-based algorithm for discovering clusters in large spatial databases with noise. In Proceedings of the 2nd International Conference on Knowledge Discovery and Data Mining, 226--231.]]"},{"key":"e_1_2_1_16_1","volume-title":"Proceedings of the ACM SIGMOD Conference, 163--174","author":"Faloutsos C.","unstructured":"Faloutsos , C. and Lin , K . -I. D. 1995. FastMap: A fast algorithm for indexing, data-mining and visualization of traditional and multimedia datasets . In Proceedings of the ACM SIGMOD Conference, 163--174 .]] 10.1145\/223784.223812 Faloutsos, C. and Lin, K.-I. D. 1995. FastMap: A fast algorithm for indexing, data-mining and visualization of traditional and multimedia datasets. In Proceedings of the ACM SIGMOD Conference, 163--174.]] 10.1145\/223784.223812"},{"key":"e_1_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1145\/355744.355745"},{"key":"e_1_2_1_18_1","volume-title":"Geometric Methods and Applications for Computer Science and Engineering","author":"Gallier J.","unstructured":"Gallier , J. 2000. Geometric Methods and Applications for Computer Science and Engineering . Springer-Verlag , New York .]] Gallier, J. 2000. Geometric Methods and Applications for Computer Science and Engineering. Springer-Verlag, New York.]]"},{"key":"e_1_2_1_19_1","volume-title":"Proceedings of the Sigmod Conference 1999, ACM Turing Award Lecture (video). ACM SIGMOD Digital Symposium Collection 2, 2.]]","author":"Gray J.","year":"2000","unstructured":"Gray , J. 2000 . What next&quest; A few remaining problems in information technlogy . In Proceedings of the Sigmod Conference 1999, ACM Turing Award Lecture (video). ACM SIGMOD Digital Symposium Collection 2, 2.]] Gray, J. 2000. What next&quest; A few remaining problems in information technlogy. In Proceedings of the Sigmod Conference 1999, ACM Turing Award Lecture (video). ACM SIGMOD Digital Symposium Collection 2, 2.]]"},{"key":"e_1_2_1_20_1","volume-title":"Proceedings of the ACM SIGMOD Conference, 73--84","author":"Guha S.","unstructured":"Guha , S. , Rastogi , R. , and Shim , K . 1998. CURE: An efficient clustering algorithm for large databases . In Proceedings of the ACM SIGMOD Conference, 73--84 .]] 10.1145\/276304.276312 Guha, S., Rastogi, R., and Shim, K. 1998. CURE: An efficient clustering algorithm for large databases. In Proceedings of the ACM SIGMOD Conference, 73--84.]] 10.1145\/276304.276312"},{"key":"e_1_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1016\/S0306-4379(00)00022-3"},{"key":"e_1_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1145\/565117.565124"},{"key":"e_1_2_1_23_1","volume-title":"Proceedings of the ACM SIGKDD Conference, 58--65","author":"Hinneburg A.","unstructured":"Hinneburg , A. and Keim , D. A . 1998. An efficient approach to clustering in large multimedia databases with noise . In Proceedings of the ACM SIGKDD Conference, 58--65 .]] Hinneburg, A. and Keim, D. A. 1998. An efficient approach to clustering in large multimedia databases with noise. In Proceedings of the ACM SIGKDD Conference, 58--65.]]"},{"key":"e_1_2_1_24_1","unstructured":"Hinneburg A. Keim D. A. and Wawryniuk M. 1999. Visual mining of high-dimensional data. IEEE Comput. Graphics Appl. 1--8.]] 10.1109\/38.788795 Hinneburg A. Keim D. A. and Wawryniuk M. 1999. Visual mining of high-dimensional data. IEEE Comput. Graphics Appl. 1--8.]] 10.1109\/38.788795"},{"key":"e_1_2_1_25_1","unstructured":"Hoffman P. Grinstein G. Marx K. Grosse I. and Stanley E. 1997. DNA visual and analytic data mining. IEEE Visualization 437--442.]] Hoffman P. Grinstein G. Marx K. Grosse I. and Stanley E. 1997. DNA visual and analytic data mining. IEEE Visualization 437--442.]]"},{"key":"e_1_2_1_26_1","volume-title":"Proceedings of the IEEE Symposium on Information Visualization, 100--107","author":"Inselberg A.","year":"1997","unstructured":"Inselberg , A. 1997 . Multidimensional detective . In Proceedings of the IEEE Symposium on Information Visualization, 100--107 .]] Inselberg, A. 1997. Multidimensional detective. In Proceedings of the IEEE Symposium on Information Visualization, 100--107.]]"},{"key":"e_1_2_1_27_1","unstructured":"Jain A. K. and Dubes R. C. 1988. Algorithms for Clustering Data. Prentice Hall New York.]] Jain A. K. and Dubes R. C. 1988. Algorithms for Clustering Data. Prentice Hall New York.]]"},{"key":"e_1_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1145\/331499.331504"},{"key":"e_1_2_1_29_1","volume-title":"Proceedings of the ACM SIGKDD Conference, 107--116","author":"Kandogan E.","year":"2001","unstructured":"Kandogan , E. 2001 . Visualizing multidimensional clusters, trends, and outliers using star coordinates . In Proceedings of the ACM SIGKDD Conference, 107--116 .]] 10.1145\/502512.502530 Kandogan, E. 2001. Visualizing multidimensional clusters, trends, and outliers using star coordinates. In Proceedings of the ACM SIGKDD Conference, 107--116.]] 10.1145\/502512.502530"},{"key":"e_1_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1109\/2.781637"},{"key":"e_1_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.1145\/381641.381656"},{"key":"e_1_2_1_32_1","doi-asserted-by":"crossref","first-page":"65","DOI":"10.1111\/j.1551-6708.1987.tb00863.x","article-title":"Why a diagram is (sometimes) worth ten thousand words","volume":"11","author":"Larkin J.","year":"1987","unstructured":"Larkin , J. and Simon , H. 1987 . Why a diagram is (sometimes) worth ten thousand words . Cognitive Sci. 11 , 65 -- 99 .]] Larkin, J. and Simon, H. 1987. Why a diagram is (sometimes) worth ten thousand words. Cognitive Sci. 11, 65--99.]]","journal-title":"Cognitive Sci."},{"key":"e_1_2_1_33_1","volume-title":"Proceedings of the IEEE Conference on Visualization, 230--239","author":"LeBlanc J.","unstructured":"LeBlanc , J. , Ward , M. O. , and Wittels , N . 1990. Exploring n-dimensional databases . In Proceedings of the IEEE Conference on Visualization, 230--239 .]] LeBlanc, J., Ward, M. O., and Wittels, N. 1990. Exploring n-dimensional databases. In Proceedings of the IEEE Conference on Visualization, 230--239.]]"},{"key":"e_1_2_1_34_1","doi-asserted-by":"publisher","DOI":"10.1109\/TKDE.2002.1019214"},{"key":"e_1_2_1_35_1","volume-title":"Using the GLYPH concept to create user-definable display formats. National Comput","author":"Littlefield R. J.","unstructured":"Littlefield , R. J. 1983. Using the GLYPH concept to create user-definable display formats. National Comput . Graphics Association , 697--706.]] Littlefield, R. J. 1983. Using the GLYPH concept to create user-definable display formats. National Comput. Graphics Association, 697--706.]]"},{"key":"e_1_2_1_36_1","doi-asserted-by":"crossref","unstructured":"Liu H. and Motoda H. 1998. Feature Extraction Construction and Selection: A Data Mining Perspective. Kluwer Academic Boston MA.]] Liu H. and Motoda H. 1998. Feature Extraction Construction and Selection: A Data Mining Perspective. Kluwer Academic Boston MA.]]","DOI":"10.1007\/978-1-4615-5725-8"},{"key":"e_1_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.1162\/153244302760200678"},{"key":"e_1_2_1_38_1","series-title":"SIAM Rev., 167--256.]]","volume-title":"The structure and function of complex networks","author":"Newman M. E.","unstructured":"Newman , M. E. 2003. The structure and function of complex networks . SIAM Rev., 167--256.]] Newman, M. E. 2003. The structure and function of complex networks. SIAM Rev., 167--256.]]"},{"key":"e_1_2_1_39_1","volume-title":"Proceedings of the ACM SIGMOD Conference, 82--92","author":"Palmer C.","year":"2009","unstructured":"Palmer , C. and Faloutsos , C . 2000. Density biased sampling: An improved method for data mining and clustering . In Proceedings of the ACM SIGMOD Conference, 82--92 .]] 10.1145\/34 2009 .335384 Palmer, C. and Faloutsos, C. 2000. Density biased sampling: An improved method for data mining and clustering. In Proceedings of the ACM SIGMOD Conference, 82--92.]] 10.1145\/342009.335384"},{"key":"e_1_2_1_40_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2005.101"},{"key":"e_1_2_1_41_1","volume-title":"Introduction to Probability Models","author":"Ross S. M.","unstructured":"Ross , S. M. 2000. Introduction to Probability Models . Academic Press , San Diego, CA .]] Ross, S. M. 2000. Introduction to Probability Models. Academic Press, San Diego, CA.]]"},{"key":"e_1_2_1_42_1","volume-title":"Proceedings of the ACM SIGIR Conference, 289--290","author":"Roussinov D.","unstructured":"Roussinov , D. , Tolle , K. , Ramsey , M. , and Chen , H . 1999. Interactive Internet search through automatic clustering: An empirical study . In Proceedings of the ACM SIGIR Conference, 289--290 .]] 10.1145\/312624.312714 Roussinov, D., Tolle, K., Ramsey, M., and Chen, H. 1999. Interactive Internet search through automatic clustering: An empirical study. In Proceedings of the ACM SIGIR Conference, 289--290.]] 10.1145\/312624.312714"},{"key":"e_1_2_1_43_1","volume-title":"Proceedings of the ACM SIGIR Conference, 74--81","author":"Schutze H.","unstructured":"Schutze , H. and Silverstein , C . 1997. Projections for efficient document clustering . In Proceedings of the ACM SIGIR Conference, 74--81 .]] 10.1145\/258525.258539 Schutze, H. and Silverstein, C. 1997. Projections for efficient document clustering. In Proceedings of the ACM SIGIR Conference, 74--81.]] 10.1145\/258525.258539"},{"key":"e_1_2_1_44_1","doi-asserted-by":"publisher","DOI":"10.1109\/MC.2002.1016905"},{"key":"e_1_2_1_45_1","volume-title":"Applied Multivariate Techniques","author":"Sharma S.","unstructured":"Sharma , S. 1995. Applied Multivariate Techniques . Wiley and Sons , New York .]] Sharma, S. 1995. Applied Multivariate Techniques. Wiley and Sons, New York.]]"},{"key":"e_1_2_1_46_1","volume-title":"Proceedings of the Very Large Databases Conference (VLDB), 428--439","author":"Sheikholeslami G.","unstructured":"Sheikholeslami , G. , Chatterjee , S. , and Zhang , A . 1998. Wavecluster: A multiresolution clustering approach for very large spatial databases . In Proceedings of the Very Large Databases Conference (VLDB), 428--439 .]] Sheikholeslami, G., Chatterjee, S., and Zhang, A. 1998. Wavecluster: A multiresolution clustering approach for very large spatial databases. In Proceedings of the Very Large Databases Conference (VLDB), 428--439.]]"},{"key":"e_1_2_1_47_1","doi-asserted-by":"publisher","DOI":"10.1057\/palgrave\/ivs\/9500006"},{"key":"e_1_2_1_48_1","volume-title":"Proceedings of the ACM SIGIR Conference, 60--66","author":"Silverstein C.","unstructured":"Silverstein , C. and Pedersen , J. O . 1997. Almost-Constant-Time clustering of arbitrary corpus subsets . In Proceedings of the ACM SIGIR Conference, 60--66 .]] 10.1145\/258525.258535 Silverstein, C. and Pedersen, J. O. 1997. Almost-Constant-Time clustering of arbitrary corpus subsets. In Proceedings of the ACM SIGIR Conference, 60--66.]] 10.1145\/258525.258535"},{"key":"e_1_2_1_49_1","unstructured":"Sonka M. Hlavac V. and Boyle R. 1999. Image Processing Analysis and Machine Vision. Brooks\/Cole Publishing Pacific Grove CA.]] Sonka M. Hlavac V. and Boyle R. 1999. Image Processing Analysis and Machine Vision. Brooks\/Cole Publishing Pacific Grove CA.]]"},{"key":"e_1_2_1_50_1","doi-asserted-by":"publisher","DOI":"10.1145\/3147.3165"},{"key":"e_1_2_1_51_1","volume-title":"Xmdvtool: Integrating multiple methods for visualizing multivariate data","author":"Ward M.","year":"1994","unstructured":"Ward , M. 1994 . Xmdvtool: Integrating multiple methods for visualizing multivariate data . IEEE Visualization , 326--333.]] Ward, M. 1994. Xmdvtool: Integrating multiple methods for visualizing multivariate data. IEEE Visualization, 326--333.]]"},{"key":"e_1_2_1_52_1","doi-asserted-by":"publisher","DOI":"10.1016\/0306-4573(88)90027-1"},{"key":"e_1_2_1_53_1","volume-title":"Proceedings of the IEEE International Conference on Data Engineering (ICDE), 324--331","author":"Xu X.","unstructured":"Xu , X. , Ester , M. , Kriegel , H.-P. , and Sander , J . 1998. A distribution-based clustering algorithm for mining in large spatial databases . In Proceedings of the IEEE International Conference on Data Engineering (ICDE), 324--331 .]] Xu, X., Ester, M., Kriegel, H.-P., and Sander, J. 1998. A distribution-based clustering algorithm for mining in large spatial databases. In Proceedings of the IEEE International Conference on Data Engineering (ICDE), 324--331.]]"},{"key":"e_1_2_1_54_1","doi-asserted-by":"crossref","first-page":"265","DOI":"10.1016\/S0097-8493(02)00283-2","article-title":"Interactive hierarchical displays: A general framework for visualization and exploration of large multivariate datasets","volume":"27","author":"Yang J.","year":"2002","unstructured":"Yang , J. , Ward , M. O. , and Rundensteiner , E. A. 2002 . Interactive hierarchical displays: A general framework for visualization and exploration of large multivariate datasets . Comput. Graphics J. 27 , 265 -- 283 .]] Yang, J., Ward, M. O., and Rundensteiner, E. A. 2002. Interactive hierarchical displays: A general framework for visualization and exploration of large multivariate datasets. Comput. Graphics J. 27, 265--283.]]","journal-title":"Comput. Graphics J."},{"key":"e_1_2_1_55_1","volume-title":"Proceedings of the ACM SIGKDD Conference, 236--243","author":"Yang L.","year":"2000","unstructured":"Yang , L. 2000 . Interactive exploration of very large relational datasets through 3D dynamic projections . In Proceedings of the ACM SIGKDD Conference, 236--243 .]] 10.1145\/347090.347134 Yang, L. 2000. Interactive exploration of very large relational datasets through 3D dynamic projections. In Proceedings of the ACM SIGKDD Conference, 236--243.]] 10.1145\/347090.347134"},{"key":"e_1_2_1_56_1","volume-title":"Proceedings of the ICML-97","author":"Yang Y.","unstructured":"Yang , Y. and Pedersen , J. O . 1997. A comparative study on feature selection in text categorization . In Proceedings of the ICML-97 , 14th International Conference on Machine Learning, 412--420.]] Yang, Y. and Pedersen, J. O. 1997. A comparative study on feature selection in text categorization. In Proceedings of the ICML-97, 14th International Conference on Machine Learning, 412--420.]]"},{"key":"e_1_2_1_57_1","volume-title":"Proceedings of the ACM SIGIR Conference, 46--54","author":"Zamir O.","unstructured":"Zamir , O. and Etzioni , O . 1998. Web document clustering: A feasibility demonstration . In Proceedings of the ACM SIGIR Conference, 46--54 .]] 10.1145\/290941.290956 Zamir, O. and Etzioni, O. 1998. Web document clustering: A feasibility demonstration. In Proceedings of the ACM SIGIR Conference, 46--54.]] 10.1145\/290941.290956"},{"key":"e_1_2_1_58_1","volume-title":"Proceedings of the ACM SIGMOD Conference, 103--114","author":"Zhang T.","unstructured":"Zhang , T. , Ramakrishnan , R. , and Livny , M . 1996. BIRCH: An efficient data clustering method for very large databases . In Proceedings of the ACM SIGMOD Conference, 103--114 .]] 10.1145\/233269.233324 Zhang, T., Ramakrishnan, R., and Livny, M. 1996. BIRCH: An efficient data clustering method for very large databases. In Proceedings of the ACM SIGMOD Conference, 103--114.]] 10.1145\/233269.233324"}],"container-title":["ACM Transactions on Information Systems"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/1148020.1148024","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/1148020.1148024","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T23:43:53Z","timestamp":1750290233000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/1148020.1148024"}},"subtitle":["Interactive visualization-based framework for clustering large datasets"],"short-title":[],"issued":{"date-parts":[[2006,4]]},"references-count":58,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2006,4]]}},"alternative-id":["10.1145\/1148020.1148024"],"URL":"https:\/\/doi.org\/10.1145\/1148020.1148024","relation":{},"ISSN":["1046-8188","1558-2868"],"issn-type":[{"value":"1046-8188","type":"print"},{"value":"1558-2868","type":"electronic"}],"subject":[],"published":{"date-parts":[[2006,4]]},"assertion":[{"value":"2006-04-01","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}