{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,9]],"date-time":"2026-07-09T16:18:37Z","timestamp":1783613917133,"version":"3.55.0"},"reference-count":18,"publisher":"Institute of Electronics, Information and Communications Engineers (IEICE)","issue":"12","content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["IEICE Trans. Inf. &amp; Syst."],"published-print":{"date-parts":[[2016]]},"DOI":"10.1587\/transinf.2016edp7130","type":"journal-article","created":{"date-parts":[[2016,11,30]],"date-time":"2016-11-30T22:14:40Z","timestamp":1480544080000},"page":"3101-3109","source":"Crossref","is-referenced-by-count":19,"title":["Cluster-Based Minority Over-Sampling for Imbalanced Datasets"],"prefix":"10.1587","volume":"E99.D","author":[{"given":"Kamthorn","family":"PUNTUMAPON","sequence":"first","affiliation":[{"name":"Department of Computer Engineering, Faculty of Engineering, Kasetsart University"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Thanawin","family":"RAKTHAMAMON","sequence":"additional","affiliation":[{"name":"Department of Computer Engineering, Faculty of Engineering, Kasetsart University"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Kitsana","family":"WAIYAMAI","sequence":"additional","affiliation":[{"name":"Department of Computer Engineering, Faculty of Engineering, Kasetsart University"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"532","reference":[{"key":"1","doi-asserted-by":"crossref","unstructured":"[1] Q. Yang and X. Wu, \u201c10 challenging problems in data mining research.,\u201d International Journal of Information Technology and Decision Making, vol.5, no.4, pp.597-604, 2006.","DOI":"10.1142\/S0219622006002258"},{"key":"2","doi-asserted-by":"crossref","unstructured":"[2] N.V. Chawla, K.W. Bowyer, L.O. Hall, and W.P. Kegelmeyer, \u201cSmote: Synthetic minority over-sampling technique,\u201d J. Artificial Intelligence Research, vol.16, pp.321-357, 2002.","DOI":"10.1613\/jair.953"},{"key":"3","doi-asserted-by":"crossref","unstructured":"[3] H. Han, W. Wang, and B. Mao, \u201cBorderline-smote: A new over-sampling method in imbalanced data sets learning,\u201d ICIC (1), pp.878-887, 2005.","DOI":"10.1007\/11538059_91"},{"key":"4","doi-asserted-by":"crossref","unstructured":"[4] C. Bunkhumpornpat, K. Sinapiromsaran, and C. Lursinsap, \u201cSafe-level-smote: Safe-level-synthetic minority over-sampling technique for handling the class imbalanced problem.,\u201d PAKDD, ed. T. Theeramunkong, B. Kijsirikul, N. Cercone, and T.B. Ho, Lect. Notes Comput. Sci., vol.5476, pp.475-482, Springer, 2009.","DOI":"10.1007\/978-3-642-01307-2_43"},{"key":"5","unstructured":"[5] H. He, Y. Bai, E.A. Garcia, and S. Li, \u201cAdasyn: Adaptive synthetic sampling approach for imbalanced learning.,\u201d IJCNN, pp.1322-1328, IEEE, 2008."},{"key":"6","doi-asserted-by":"crossref","unstructured":"[6] S. Barua, M.M. Islam, X. Yao, and K. Murase, \u201cMwmote-majority weighted minority oversampling technique for imbalanced data set learning.,\u201d IEEE Trans. Knowl. Data Eng., vol.26, no.2, pp.405-425, 2014.","DOI":"10.1109\/TKDE.2012.232"},{"key":"7","doi-asserted-by":"crossref","unstructured":"[7] M. P\u00e9rez-Ortiz, P.A. Guti\u00e9rrez, C. Herv\u00e1s-Mart\u00ednez, and X. Yao, \u201cGraph-based approaches for over-sampling in the context of ordinal regression.,\u201d IEEE Trans. Knowl. Data Eng., vol.27, no.5, pp.1233-1245, 2015.","DOI":"10.1109\/TKDE.2014.2365780"},{"key":"8","doi-asserted-by":"crossref","unstructured":"[8] K. Puntumapon and K. Waiyamai, \u201cA pruning-based approach for searching precise and generalized region for synthetic minority over-sampling,\u201d PAKDD (2), ed. P.N. Tan, S. Chawla, C.K. Ho, and J. Bailey, Lect. Notes Comput. Sci., vol.7302, pp.371-382, Springer, 2012.","DOI":"10.1007\/978-3-642-30220-6_31"},{"key":"9","doi-asserted-by":"crossref","unstructured":"[9] X. Fan, K. Tang, and T. Weise, \u201cMargin-Based Over-Sampling Method for Learning From Imbalanced Datasets,\u201d Proc. 15th Pacific-Asia Conference on Knowledge Discovery and Data Mining (PAKDD&apos;11), ed. J.Z. Huang, L. Cao, and J. Srivastava, Lect. Notes Comput. Sci. (LNCS), pp.309-320, Springer-Verlag GmbH: Berlin, Germany, 2011.","DOI":"10.1007\/978-3-642-20847-8_26"},{"key":"10","doi-asserted-by":"crossref","unstructured":"[10] S.J. Yen and Y.S. Lee, \u201cCluster-based under-sampling approaches for imbalanced data distributions.,\u201d Expert Syst. Appl., vol.36, no.3, pp.5718-5727, 2009.","DOI":"10.1016\/j.eswa.2008.06.108"},{"key":"11","unstructured":"[11] M. Kubat and S. Matwin, \u201cAddressing the curse of imbalanced training sets: One-sided selection.,\u201d ICML, ed. D.H. Fisher, pp.179-186, Morgan Kaufmann, 1997."},{"key":"12","doi-asserted-by":"crossref","unstructured":"[12] I. Tomek, \u201cTwo Modifications of CNN,\u201d IEEE Trans. Syst. Man Cybern., vol.6, no.11, pp.769-772, 1976.","DOI":"10.1109\/TSMC.1976.4309452"},{"key":"13","unstructured":"[13] J. Zhang and I. Mani, \u201cKNN Approach to Unbalanced Data Distributions: A Case Study Involving Information Extraction,\u201d Proc. ICML&apos;2003 Workshop on Learning from Imbalanced Datasets, 2003."},{"key":"14","doi-asserted-by":"crossref","unstructured":"[14] A. Bradley, \u201cThe use of the area under the ROC curve in the evaluation of machine learning algorithms,\u201d Pattern Recognit., vol.30, no.7, pp.1145-1159, 1997.","DOI":"10.1016\/S0031-3203(96)00142-2"},{"key":"15","doi-asserted-by":"crossref","unstructured":"[15] C.J. van Rijsbergen, Information Retrieval, 2nd ed., Butterworths, London, 1979.","DOI":"10.1007\/978-3-642-23318-0_2"},{"key":"16","doi-asserted-by":"crossref","unstructured":"[16] M. Hall, E. Frank, G. Holmes, B. Pfahringer, P. Reutemann, and I.H. Witten, \u201cThe WEKA data mining software: An update,\u201d SIGKDD Explorations, vol.11, no.1, pp.10-18, 2009.","DOI":"10.1145\/1656274.1656278"},{"key":"17","unstructured":"[17] A. Frank and A. Asuncion, \u201cUCI machine learning repository,\u201d 2010."},{"key":"18","doi-asserted-by":"crossref","unstructured":"[18] P.H. Ramsey, J.L. Hodges, and J.P. Shaffer, \u201cSignificance probabilities of the wilcoxon signed-rank test,\u201d J. Nonparametric Statistics, vol.2, no.2, pp.133-153, 1993.","DOI":"10.1080\/10485259308832548"}],"container-title":["IEICE Transactions on Information and Systems"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.jstage.jst.go.jp\/article\/transinf\/E99.D\/12\/E99.D_2016EDP7130\/_pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2023,8,21]],"date-time":"2023-08-21T04:00:11Z","timestamp":1692590411000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.jstage.jst.go.jp\/article\/transinf\/E99.D\/12\/E99.D_2016EDP7130\/_article"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2016]]},"references-count":18,"journal-issue":{"issue":"12","published-print":{"date-parts":[[2016]]}},"URL":"https:\/\/doi.org\/10.1587\/transinf.2016edp7130","relation":{},"ISSN":["0916-8532","1745-1361"],"issn-type":[{"value":"0916-8532","type":"print"},{"value":"1745-1361","type":"electronic"}],"subject":[],"published":{"date-parts":[[2016]]}}}