{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,9,4]],"date-time":"2025-09-04T13:57:47Z","timestamp":1756994267178,"version":"3.37.3"},"reference-count":36,"publisher":"Springer Science and Business Media LLC","issue":"1","license":[{"start":{"date-parts":[[2016,2,6]],"date-time":"2016-02-06T00:00:00Z","timestamp":1454716800000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2016,2,6]],"date-time":"2016-02-06T00:00:00Z","timestamp":1454716800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["BMC Bioinformatics"],"abstract":"<jats:title>Abstract<\/jats:title><jats:sec>\n                <jats:title>Background<\/jats:title>\n                <jats:p>Gene set analysis (GSA) aims to evaluate the association between the expression of biological pathways, or a priori defined gene sets, and a particular phenotype. Numerous GSA methods have been proposed to assess the enrichment of sets of genes. However, most methods are developed with respect to a specific alternative scenario, such as a differential mean pattern or a differential coexpression. Moreover, a very limited number of methods can handle either binary, categorical, or continuous phenotypes. In this paper, we develop two novel GSA tests, called SDRs, based on the sufficient dimension reduction technique, which aims to capture sufficient information about the relationship between genes and the phenotype. The advantages of our proposed methods are that they allow for categorical and continuous phenotypes, and they are also able to identify a variety of enriched gene sets.<\/jats:p>\n              <\/jats:sec><jats:sec>\n                <jats:title>Results<\/jats:title>\n                <jats:p>Through simulation studies, we compared the type I error and power of SDRs with existing GSA methods for binary, triple, and continuous phenotypes. We found that SDR methods adequately control the type I error rate at the pre-specified nominal level, and they have a satisfactory power to detect gene sets with differential coexpression and to test non-linear associations between gene sets and a continuous phenotype. In addition, the SDR methods were compared with seven widely-used GSA methods using two real microarray datasets for illustration.<\/jats:p>\n              <\/jats:sec><jats:sec>\n                <jats:title>Conclusions<\/jats:title>\n                <jats:p>We concluded that the SDR methods outperform the others because of their flexibility with regard to handling different kinds of phenotypes and their power to detect a wide range of alternative scenarios. Our real data analysis highlights the differences between GSA methods for detecting enriched gene sets.<\/jats:p>\n              <\/jats:sec>","DOI":"10.1186\/s12859-016-0928-6","type":"journal-article","created":{"date-parts":[[2016,2,6]],"date-time":"2016-02-06T04:46:55Z","timestamp":1454734015000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":8,"title":["Gene set analysis using sufficient dimension reduction"],"prefix":"10.1186","volume":"17","author":[{"given":"Huey-Miin","family":"Hsueh","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Chen-An","family":"Tsai","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2016,2,6]]},"reference":[{"issue":"8","key":"928_CR1","doi-asserted-by":"publisher","first-page":"980","DOI":"10.1093\/bioinformatics\/btm051","volume":"23","author":"JJ Goeman","year":"2007","unstructured":"Goeman JJ, B\u00fchmann P. Analyzing gene expression data in terms of gene sets: methodological issues. Bioinformatics. 2007; 23(8):980\u20137.","journal-title":"Bioinformatics"},{"key":"928_CR2","doi-asserted-by":"publisher","first-page":"189","DOI":"10.1093\/bib\/bbn001","volume":"9","author":"D Nam","year":"2008","unstructured":"Nam D, Kim SY. Gene-set approach for expression pattern analysis. Brief Bioinform. 2008; 9:189\u201397.","journal-title":"Brief Bioinform"},{"issue":"1","key":"928_CR3","doi-asserted-by":"publisher","first-page":"24","DOI":"10.1093\/bib\/bbn042","volume":"10","author":"I Dinu","year":"2008","unstructured":"Dinu I, Potter JD, Mueller T, Liu Q, Adewale AJ, Jhangri GS, et al.Gene-set analysis and reduction. Brief Bioinform. 2008; 10(1):24\u201334.","journal-title":"Brief Bioinform"},{"issue":"4","key":"928_CR4","doi-asserted-by":"publisher","first-page":"504","DOI":"10.1093\/bib\/bbt002","volume":"15","author":"H Maciejewski","year":"2014","unstructured":"Maciejewski H. Gene set analysis methods: statistical models and methodological differences. Brief Bioinform. 2014; 15(4):504\u201318.","journal-title":"Brief Bioinform"},{"issue":"43","key":"928_CR5","doi-asserted-by":"publisher","first-page":"15545","DOI":"10.1073\/pnas.0506580102","volume":"102","author":"A Subramanian","year":"2005","unstructured":"Subramanian A, Tamayo P, Mootha VK, Mhkherjee S, Ebert BL, Gillette MA, et al.Gene set enrichment analysis: a knowledge-based approach for interpreting genome-wide expression profiles. Proc Natl Acad Sci U S A. 2005; 102(43):15545\u201350.","journal-title":"Proc Natl Acad Sci U S A"},{"issue":"38","key":"928_CR6","doi-asserted-by":"publisher","first-page":"13544","DOI":"10.1073\/pnas.0506577102","volume":"102","author":"L Tian","year":"2005","unstructured":"Tian L, Greenberg SA, Kong SW, Altschuler J, Kohane I, Park PJ. Discovering statistically significant pathways in expression profiling studies. Proc Natl Acad Sci U S A. 2005; 102(38):13544\u20139.","journal-title":"Proc Natl Acad Sci U S A"},{"issue":"1","key":"928_CR7","doi-asserted-by":"publisher","first-page":"107","DOI":"10.1214\/07-AOAS101","volume":"1","author":"B Efron","year":"2007","unstructured":"Efron B, Tibshirani R. On testing the significance of sets of genes. Ann Appl Stat. 2007; 1(1):107\u201329.","journal-title":"Ann Appl Stat."},{"issue":"6","key":"928_CR8","doi-asserted-by":"publisher","first-page":"565","DOI":"10.1177\/0962280209351908","volume":"18","author":"RA Irizarry","year":"2009","unstructured":"Irizarry RA, Wang C, Zhou Y, Speed TP. Gene set enrichment analysis made simple. Stat Methods Med Res. 2009; 18(6):565\u201375.","journal-title":"Stat Methods Med Res"},{"issue":"3","key":"928_CR9","doi-asserted-by":"publisher","first-page":"306","DOI":"10.1093\/bioinformatics\/btl599","volume":"23","author":"Y Jiang","year":"2007","unstructured":"Jiang Y, Gentleman R. Extensions to gene set enrichment. Bioinformatics. 2007; 23(3):306\u201313.","journal-title":"Bioinformatics"},{"issue":"19","key":"928_CR10","doi-asserted-by":"publisher","first-page":"2373","DOI":"10.1093\/bioinformatics\/btl401","volume":"22","author":"SW Kong","year":"2006","unstructured":"Kong SW, Pu WT, Park PJ. A multivariate approach for integrating genome-wide expression data and biological knowledge. Bioinformatics. 2006; 22(19):2373\u201380.","journal-title":"Bioinformatics"},{"issue":"7","key":"928_CR11","doi-asserted-by":"publisher","first-page":"897","DOI":"10.1093\/bioinformatics\/btp098","volume":"25","author":"CA Tsai","year":"2009","unstructured":"Tsai CA, Chen JJ. Bioinformatics. 2009; 25(7):897\u2013903.","journal-title":"Bioinformatics"},{"key":"928_CR12","doi-asserted-by":"crossref","unstructured":"Chien CY, Chang CW, Tsai CA, Chen JJ. MAVTgsa: An R package for gene set (enrichment) analysis. BioMed Res Int. 2014;2014(346074). doi:http:\/\/dx.doi.org\/10.1155\/2014\/346074.","DOI":"10.1155\/2014\/346074"},{"key":"928_CR13","doi-asserted-by":"publisher","first-page":"249","DOI":"10.1126\/science.1087447","volume":"302","author":"JM Stuart","year":"2003","unstructured":"Stuart JM, Segal E, Koller D, Kim SK. A gene-coexpression network for global discovery of conserved genetic modules. Science. 2003; 302:249\u201354.","journal-title":"Science"},{"key":"928_CR14","doi-asserted-by":"crossref","first-page":"Article 17","DOI":"10.2202\/1544-6115.1128","volume":"4","author":"B Zhang","year":"2005","unstructured":"Zhang B, Horvath S. A general framework for weighted gene co-expression network analysis. Stat Appl Genet Mol Biol. 2005; 4:Article 17.","journal-title":"Stat Appl Genet Mol Biol"},{"key":"928_CR15","doi-asserted-by":"publisher","first-page":"109","DOI":"10.1186\/1471-2105-10-109","volume":"10","author":"SB Cho","year":"2009","unstructured":"Cho SB, Kim J, Kim JH. Identifying set-wise differential co-expression in gene expression microarray data. BMC Bioinformatics. 2009; 10:109.","journal-title":"BMC Bioinformatics"},{"issue":"24","key":"928_CR16","doi-asserted-by":"publisher","first-page":"4348","DOI":"10.1093\/bioinformatics\/bti722","volume":"21","author":"JK Choi","year":"2005","unstructured":"Choi JK, Yu U, Yoo OJ, Kim S. Differential coexpression analysis using microarray data and its application to human cancer. Bioinformatics. 2005; 21(24):4348\u201355.","journal-title":"Bioinformatics"},{"issue":"21","key":"928_CR17","doi-asserted-by":"publisher","first-page":"2780","DOI":"10.1093\/bioinformatics\/btp502","volume":"25","author":"YJ Choi","year":"2009","unstructured":"Choi YJ, Kendziorski C. Statistical methods for gene set co-expression analysis. Bioinformatics. 2009; 25(21):2780\u20136.","journal-title":"Bioinformatics"},{"issue":"3","key":"928_CR18","doi-asserted-by":"publisher","first-page":"360","DOI":"10.1093\/bioinformatics\/btt687","volume":"30","author":"Y Rahmatallah","year":"2014","unstructured":"Rahmatallah Y, Emmert-Streib F, Glazko G. Gene sets net correlations analysis (GSNCA): a multivariate differential coexpression test for gene sets. Bioinformatics. 2014; 30(3):360\u20138.","journal-title":"Bioinformatics"},{"issue":"7","key":"928_CR19","doi-asserted-by":"publisher","first-page":"e60","DOI":"10.1093\/nar\/gku099","volume":"42","author":"S Jung","year":"2014","unstructured":"Jung S, Kim S. EDDY: a novel statistical gene set test method to detect differential genetic dependencies. Nucleic Acid Res. 2014; 42(7):e60.","journal-title":"Nucleic Acid Res"},{"issue":"23","key":"928_CR20","doi-asserted-by":"publisher","first-page":"3073","DOI":"10.1093\/bioinformatics\/bts579","volume":"28","author":"Y Rahmatallah","year":"2013","unstructured":"Rahmatallah Y, Emmert-Streib F, Glazko G. Gene set analysis for self-contained tests: complex null and specific alternative hypotheses. Bioinformatics. 2013; 28(23):3073\u201380.","journal-title":"Bioinformatics"},{"issue":"1","key":"928_CR21","doi-asserted-by":"publisher","first-page":"93","DOI":"10.1093\/bioinformatics\/btg382","volume":"20","author":"JJ Goeman","year":"2004","unstructured":"Goeman JJ, van de Geer S, de Kort F, van Houwelingen HC. A global test for groups of genes: testing association with a clinical outcome. Bioinformatics. 2004; 20(1):93\u20139.","journal-title":"Bioinformatics"},{"key":"928_CR22","doi-asserted-by":"publisher","first-page":"212","DOI":"10.1186\/1471-2105-14-212","volume":"14","author":"I Dinu","year":"2013","unstructured":"Dinu I, Wang X, Kelemen LE, Vatanpour S, Pyne S. Linear combination test for gene set analysis of a continuous phenotype. BMC Bioinformatics. 2013; 14:212.","journal-title":"BMC Bioinformatics"},{"key":"928_CR23","doi-asserted-by":"publisher","first-page":"260","DOI":"10.1186\/1471-2105-15-260","volume":"15","author":"X Wang","year":"2014","unstructured":"Wang X, Pyne S, Dinu I. Gene set enrichment analysis for multiple continuous phenotypes. BMC Bioinformatics. 2014; 15:260.","journal-title":"BMC Bioinformatics"},{"issue":"414","key":"928_CR24","doi-asserted-by":"publisher","first-page":"316","DOI":"10.1080\/01621459.1991.10475035","volume":"86","author":"KC Li","year":"1991","unstructured":"Li KC. Sliced inverse regression for dimension reduction. J Am Stat Assoc. 1991; 86(414):316\u201327.","journal-title":"J Am Stat Assoc"},{"key":"928_CR25","doi-asserted-by":"publisher","first-page":"130","DOI":"10.1016\/j.jmva.2010.08.007","volume":"102","author":"E Bura","year":"2011","unstructured":"Bura E, Yang J. Dimension estimation in sufficient dimension reduction: a unifying approach. J Multivar Anal. 2011; 102:130\u201342.","journal-title":"J Multivar Anal"},{"issue":"414","key":"928_CR26","first-page":"328","volume":"86","author":"RD Cook","year":"1991","unstructured":"Cook RD, Weisberg S. Discussion of \u201cSliced inverse regression for dimension reduction\u2019. J Am Stat Assoc. 1991; 86(414):328\u201332.","journal-title":"J Am Stat Assoc."},{"issue":"448","key":"928_CR27","doi-asserted-by":"publisher","first-page":"1187","DOI":"10.1080\/01621459.1999.10473873","volume":"84","author":"RD Cook","year":"1999","unstructured":"Cook RD, Lee H. Dimension reduction in regressions with a binary response. J Am Stat Assoc. 1999; 84(448):1187\u2013200.","journal-title":"J Am Stat Assoc"},{"key":"928_CR28","doi-asserted-by":"publisher","first-page":"285","DOI":"10.1093\/biomet\/asm021","volume":"94","author":"Y Shao","year":"2007","unstructured":"Shao Y, Cook RD, Weisberg S. Marginal tests with sliced average variance estimation. Biometrika. 2007; 94:285\u201396.","journal-title":"Biometrika"},{"key":"928_CR29","doi-asserted-by":"publisher","first-page":"242","DOI":"10.1186\/1471-2105-8-242","volume":"8","author":"I Dinu","year":"2007","unstructured":"Dinu I, Potter JD, Mueller T, Liu Q, Adewale AJ, Jhangri GS, et al.Improving gene set analysis of microarray data by SAM-GS. BMC Bioinformatics. 2007; 8:242.","journal-title":"BMC Bioinformatics"},{"issue":"3","key":"928_CR30","doi-asserted-by":"publisher","first-page":"927","DOI":"10.1158\/0008-5472.CAN-07-2608","volume":"68","author":"TA Wallace","year":"2008","unstructured":"Wallace TA, Prueitt RL, Yi M, Howe TM, Gillespie JW, Yfantis HG, et al.Tumor immunobiological differences in prostate cancer between African-American and European-American men. Cancer Res. 2008; 68(3):927\u201336.","journal-title":"Cancer Res"},{"issue":"1","key":"928_CR31","doi-asserted-by":"publisher","first-page":"207","DOI":"10.1093\/nar\/30.1.207","volume":"30","author":"R Edgar","year":"2002","unstructured":"Edgar R, Domrachev M, Lash AE. Gene Expression Omnibus: NCBI gene expression and hybridization array data repository. Nucleic Acids Res. 2002; 30(1):207\u201310.","journal-title":"Nucleic Acids Res"},{"key":"928_CR32","doi-asserted-by":"publisher","first-page":"800","DOI":"10.1016\/j.eururo.2012.11.013","volume":"63","author":"EH Allott","year":"2013","unstructured":"Allott EH, Masko EM, Freedland SJ. Obesity and prostate cancer: weighing the evidence. Eur Urol. 2013; 63:800\u20139.","journal-title":"Eur Urol"},{"issue":"2","key":"928_CR33","first-page":"73","volume":"6","author":"SJ Freedland","year":"2004","unstructured":"Freedland SJ, Aronson WJ. Examining the relationship between obesity and prostate cancer. Rev Urol. 2004; 6(2):73\u201381.","journal-title":"Rev Urol"},{"key":"928_CR34","doi-asserted-by":"crossref","first-page":"Article 34","DOI":"10.2202\/1544-6115.1175","volume":"4","author":"J Sch\u00e4fer","year":"2005","unstructured":"Sch\u00e4fer J, Strimmer K. A shrinkage approach to large-scale covariance matrix estimation and implications for functional genomics. Stat Appl Genet Mol Biol. 2005; 4:Article 34.","journal-title":"Stat Appl Genet Mol Biol."},{"key":"928_CR35","unstructured":"Becker C, Gather U. A note on the choice of the number of slices in sliced inverse regression, Technical Reports. Technische Universit\u00e4t Dortmund; 2007."},{"key":"928_CR36","doi-asserted-by":"publisher","first-page":"577","DOI":"10.1177\/0962280209351925","volume":"18","author":"M Wu","year":"2009","unstructured":"Wu M, Lin X. Prior biological knowledge-based approaches for the analysis of genome-wide expression profiles using gene sets and pathways. Stat Methods Med Res. 2009; 18:577\u201393.","journal-title":"Stat Methods Med Res"}],"container-title":["BMC Bioinformatics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1186\/s12859-016-0928-6.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1186\/s12859-016-0928-6\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/link.springer.com\/content\/pdf\/10.1186\/s12859-016-0928-6","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1186\/s12859-016-0928-6.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,2,1]],"date-time":"2024-02-01T18:17:13Z","timestamp":1706811433000},"score":1,"resource":{"primary":{"URL":"https:\/\/bmcbioinformatics.biomedcentral.com\/articles\/10.1186\/s12859-016-0928-6"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2016,2,6]]},"references-count":36,"journal-issue":{"issue":"1","published-online":{"date-parts":[[2016,12]]}},"alternative-id":["928"],"URL":"https:\/\/doi.org\/10.1186\/s12859-016-0928-6","relation":{},"ISSN":["1471-2105"],"issn-type":[{"type":"electronic","value":"1471-2105"}],"subject":[],"published":{"date-parts":[[2016,2,6]]},"assertion":[{"value":"29 September 2015","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"1 February 2016","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"6 February 2016","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}}],"article-number":"74"}}