{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,19]],"date-time":"2025-06-19T04:21:39Z","timestamp":1750306899370,"version":"3.41.0"},"reference-count":31,"publisher":"Association for Computing Machinery (ACM)","issue":"3","license":[{"start":{"date-parts":[[2013,10,1]],"date-time":"2013-10-01T00:00:00Z","timestamp":1380585600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/100000144","name":"Division of Computer and Network Systems","doi-asserted-by":"publisher","award":["CNS-09-58854435060"],"award-info":[{"award-number":["CNS-09-58854435060"]}],"id":[{"id":"10.13039\/100000144","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Manage. Inf. Syst."],"published-print":{"date-parts":[[2013,10]]},"abstract":"<jats:p>When a medical practitioner encounters a patient with rare symptoms that translates to rare occurrences in the local database, it is quite valuable to draw conclusions collectively from such occurrences in other hospitals. However, for such rare conditions, there will be a huge imbalance in classes among the relevant base population. Due to regulations and privacy concerns, collecting data from other hospitals will be problematic. Consequently, distributed decision support systems that can use just the statistics of data from multiple hospitals are valuable. We present a system that can collectively build a distributed classification model dynamically without the need of patient data from each site in the case of imbalanced data. The system uses a voting ensemble of experts for the decision model. The imbalance condition and number of experts can be determined by the system. Since only statistics of the data and no raw data are required by the system, patient privacy issues are addressed. We demonstrate the outlined principles using the Nationwide Inpatient Sample (NIS) database. Results of experiments conducted on 7,810,762 patients from 1050 hospitals show improvement of 13.68% to 24.46% in balanced prediction accuracy using our model over the baseline model, illustrating the effectiveness of the proposed methodology.<\/jats:p>","DOI":"10.1145\/2517310","type":"journal-article","created":{"date-parts":[[2014,4,23]],"date-time":"2014-04-23T13:52:04Z","timestamp":1398261124000},"page":"1-15","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":6,"title":["Distributed Privacy-Preserving Decision Support System for Highly Imbalanced Clinical Data"],"prefix":"10.1145","volume":"4","author":[{"given":"George","family":"Mathew","sequence":"first","affiliation":[{"name":"Temple University"}]},{"given":"Zoran","family":"Obradovic","sequence":"additional","affiliation":[{"name":"Temple University"}]}],"member":"320","published-online":{"date-parts":[[2013,10]]},"reference":[{"key":"e_1_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/HICSS.2011.285"},{"key":"e_1_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1109\/TKDE.2005.129"},{"key":"e_1_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1145\/6592.6597"},{"key":"e_1_2_1_4_1","unstructured":"Buchanan B. G. and Shortliffe E. W. 1984. Rule Based Expert Systems: The MYCIN Experiments in the Stanford Heuristic Programming Project Addison-Wesley Reading MA.   Buchanan B. G. and Shortliffe E. W. 1984. Rule Based Expert Systems: The MYCIN Experiments in the Stanford Heuristic Programming Project Addison-Wesley Reading MA."},{"key":"e_1_2_1_5_1","first-page":"1","article-title":"A framework for learning from distributed data using sufficient statistics and its application to learning decision trees","volume":"1","author":"Caragea D.","year":"2004","unstructured":"Caragea , D. , Silvescu , A. , and Honovar , V. 2004 . A framework for learning from distributed data using sufficient statistics and its application to learning decision trees . Int. J. Hybrid Intell. Syst. 1 , 1 -- 2 , 80--89. Caragea, D., Silvescu, A., and Honovar, V. 2004. A framework for learning from distributed data using sufficient statistics and its application to learning decision trees. Int. J. Hybrid Intell. Syst. 1, 1--2, 80--89.","journal-title":"Int. J. Hybrid Intell. Syst."},{"key":"e_1_2_1_6_1","unstructured":"CCS. Clinical Classifications Software (CCS) for ICD-9-CM. Appendix A: Single-Level Diagnoses. http:\/\/www.hcup-us.ahrq.gov\/toolssoftware\/ccs\/ccs.jsp.  CCS. Clinical Classifications Software (CCS) for ICD-9-CM. Appendix A: Single-Level Diagnoses. http:\/\/www.hcup-us.ahrq.gov\/toolssoftware\/ccs\/ccs.jsp."},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/PROC.1979.11321"},{"key":"e_1_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1136\/jamia.2001.0080552"},{"key":"e_1_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1145\/508171.508174"},{"key":"e_1_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1145\/1321440.1321461"},{"key":"e_1_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/34.58871"},{"key":"e_1_2_1_12_1","volume-title":"Proceedings of the International Conference on Artificial Intelligence. 111--117","author":"Japkowicz N.","year":"2000","unstructured":"Japkowicz , N. 2000 . The class imbalance problem: Significance and strategies . In Proceedings of the International Conference on Artificial Intelligence. 111--117 . Japkowicz, N. 2000. The class imbalance problem: Significance and strategies. In Proceedings of the International Conference on Artificial Intelligence. 111--117."},{"volume-title":"Proceedings of the 3rd SIAM International Conference on Data Mining (SDM). 119--129","author":"Jin R.","key":"e_1_2_1_13_1","unstructured":"Jin , R. and Agrawal , G . 2003. Communication and memory efficient parallel decision tree construction . In Proceedings of the 3rd SIAM International Conference on Data Mining (SDM). 119--129 . Jin, R. and Agrawal, G. 2003. Communication and memory efficient parallel decision tree construction. In Proceedings of the 3rd SIAM International Conference on Data Mining (SDM). 119--129."},{"key":"e_1_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1001\/jama.2011.1515"},{"key":"e_1_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1186\/1472-6947-11-51"},{"key":"e_1_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.29012\/jpc.v1i1.566"},{"key":"e_1_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCABS.2011.5729866"},{"key":"e_1_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICMLA.2012.180"},{"key":"e_1_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1145\/356893.356898"},{"volume-title":"Proceedings of AMIA Annual Symposium. 573--577","author":"Oster S.","key":"e_1_2_1_20_1","unstructured":"Oster , S. , Langella , S. , Hastings , S. , Ervin , D. , Madduri , R. , Kurc , T. , Siebenlist , F. , Covitz , P. , Shanbhag , K. , Foster , I. , and Saltz , J . 2007 . In Proceedings of AMIA Annual Symposium. 573--577 . Oster, S., Langella, S., Hastings, S., Ervin, D., Madduri, R., Kurc, T., Siebenlist, F., Covitz, P., Shanbhag, K., Foster, I., and Saltz, J. 2007. In Proceedings of AMIA Annual Symposium. 573--577."},{"key":"e_1_2_1_21_1","article-title":"Ensemble based systems in decision making. IEEE Circuits","author":"Polikar R.","year":"2006","unstructured":"Polikar , R. 2006 . Ensemble based systems in decision making. IEEE Circuits Syst. Mag. Third Quarter, 21--45. Polikar, R. 2006. Ensemble based systems in decision making. IEEE Circuits Syst. Mag. Third Quarter, 21--45.","journal-title":"Syst. Mag. Third Quarter, 21--45."},{"volume-title":"Proceedings of the IEEE International Conference on Fuzzy Systems.","author":"Popescu M.","key":"e_1_2_1_22_1","unstructured":"Popescu , M. and Khalilia , M . 2011. Improving disease prediction using ICD-9 ontological features . In Proceedings of the IEEE International Conference on Fuzzy Systems. Popescu, M. and Khalilia, M. 2011. Improving disease prediction using ICD-9 ontological features. In Proceedings of the IEEE International Conference on Fuzzy Systems."},{"key":"e_1_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1023\/A:1022643204877"},{"key":"e_1_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1038\/35015718"},{"key":"e_1_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1136\/jamia.2001.0080527"},{"key":"e_1_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1186\/1472-6947-6-6"},{"volume-title":"Proceedings of the ACM SIGKDD Workshop on Health Informatics, in Conjunction with 18th SIGKDD Conference on Knowledge Discovery and Data Mining.","author":"Stiglic G.","key":"e_1_2_1_27_1","unstructured":"Stiglic , G. , Pernek , I. , Kokol , P. , and Obradovic , Z . 2012. Disease prediction based on prior knowledge . In Proceedings of the ACM SIGKDD Workshop on Health Informatics, in Conjunction with 18th SIGKDD Conference on Knowledge Discovery and Data Mining. Stiglic, G., Pernek, I., Kokol, P., and Obradovic, Z. 2012. Disease prediction based on prior knowledge. In Proceedings of the ACM SIGKDD Workshop on Health Informatics, in Conjunction with 18th SIGKDD Conference on Knowledge Discovery and Data Mining."},{"key":"e_1_2_1_28_1","unstructured":"Tan P. Steinbach M. and Kumar V. 2006. Introduction to Data Mining. Pearson Addison Wesley Boston MA 160.   Tan P. Steinbach M. and Kumar V. 2006. Introduction to Data Mining. Pearson Addison Wesley Boston MA 160."},{"key":"e_1_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.crad.2006.09.032"},{"key":"e_1_2_1_30_1","unstructured":"Wier L. M. Elixhauser A. Pfunter A. and Au D. H. 2011. Overview of hospitalizations among patients with COPD 2008; Statistical Brief #106. http:\/\/www.hcup-us.ahrq.gov\/reports\/statbriefs\/sb106.jsp.  Wier L. M. Elixhauser A. Pfunter A. and Au D. H. 2011. Overview of hospitalizations among patients with COPD 2008; Statistical Brief #106. http:\/\/www.hcup-us.ahrq.gov\/reports\/statbriefs\/sb106.jsp."},{"key":"e_1_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.1186\/1472-6947-10-16"}],"container-title":["ACM Transactions on Management Information Systems"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2517310","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/2517310","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T08:19:08Z","timestamp":1750234748000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2517310"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2013,10]]},"references-count":31,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2013,10]]}},"alternative-id":["10.1145\/2517310"],"URL":"https:\/\/doi.org\/10.1145\/2517310","relation":{},"ISSN":["2158-656X","2158-6578"],"issn-type":[{"type":"print","value":"2158-656X"},{"type":"electronic","value":"2158-6578"}],"subject":[],"published":{"date-parts":[[2013,10]]},"assertion":[{"value":"2012-11-01","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2013-08-01","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2013-10-01","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}