{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,7,30]],"date-time":"2025-07-30T11:43:55Z","timestamp":1753875835827,"version":"3.41.2"},"reference-count":57,"publisher":"Oxford University Press (OUP)","issue":"6","license":[{"start":{"date-parts":[[2022,9,30]],"date-time":"2022-09-30T00:00:00Z","timestamp":1664496000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/academic.oup.com\/journals\/pages\/open_access\/funder_policies\/chorus\/standard_publication_model"}],"funder":[{"DOI":"10.13039\/100000002","name":"National Institutes of Health","doi-asserted-by":"publisher","award":["R03CA245771","U10CA180821","U10CA180 882","U24CA196171","UG1CA233180","UG1CA233331","UG1CA233338","R35CA197734","P30CA016058"],"award-info":[{"award-number":["R03CA245771","U10CA180821","U10CA180 882","U24CA196171","UG1CA233180","UG1CA233331","UG1CA233338","R35CA197734","P30CA016058"]}],"id":[{"id":"10.13039\/100000002","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2022,11,19]]},"abstract":"<jats:title>Abstract<\/jats:title><jats:p>For many high-dimensional genomic and epigenomic datasets, the outcome of interest is ordinal. While these ordinal outcomes are often thought of as the observed cutpoints of some latent continuous variable, some ordinal outcomes are truly discrete and are comprised of the subjective combination of several factors. The nonlinear stereotype logistic model, which does not assume proportional odds, was developed for these \u2018assessed\u2019 ordinal variables. It has previously been extended to the frequentist high-dimensional feature selection setting, but the Bayesian framework provides some distinct advantages in terms of simultaneous uncertainty quantification and variable selection. Here, we review the stereotype model and Bayesian variable selection methods and demonstrate how to combine them to select genomic features associated with discrete ordinal outcomes. We compared the Bayesian and frequentist methods in terms of variable selection performance. We additionally applied the Bayesian stereotype method to an acute myeloid leukemia RNA-sequencing dataset to further demonstrate its variable selection abilities by identifying features associated with the European LeukemiaNet prognostic risk score.<\/jats:p>","DOI":"10.1093\/bib\/bbac414","type":"journal-article","created":{"date-parts":[[2022,10,3]],"date-time":"2022-10-03T00:59:13Z","timestamp":1664758753000},"source":"Crossref","is-referenced-by-count":1,"title":["High-dimensional genomic feature selection with the ordered stereotype logit model"],"prefix":"10.1093","volume":"23","author":[{"given":"Anna Eames","family":"Seffernick","sequence":"first","affiliation":[{"name":"Division of Biostatistics, College of Public Health, The Ohio State University , Columbus, OH, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Krzysztof","family":"Mr\u00f3zek","sequence":"additional","affiliation":[{"name":"Clara D. Bloomfield Center for Leukemia Outcomes Research, The Ohio State University , Columbus, OH, USA"},{"name":"The Ohio State Comprehensive Cancer Center , Columbus, OH, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Deedra","family":"Nicolet","sequence":"additional","affiliation":[{"name":"Clara D. Bloomfield Center for Leukemia Outcomes Research, The Ohio State University , Columbus, OH, USA"},{"name":"The Ohio State Comprehensive Cancer Center , Columbus, OH, USA"},{"name":"Alliance Statistics and Data Management Center, The Ohio State University Comprehensive Cancer Center , Columbus, OH, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Richard M","family":"Stone","sequence":"additional","affiliation":[{"name":"Dana Farber\/Partners Cancer Care, Harvard University , Boston, MA, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Ann-Kathrin","family":"Eisfeld","sequence":"additional","affiliation":[{"name":"Clara D. Bloomfield Center for Leukemia Outcomes Research, The Ohio State University , Columbus, OH, USA"},{"name":"The Ohio State Comprehensive Cancer Center , Columbus, OH, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"John C","family":"Byrd","sequence":"additional","affiliation":[{"name":"Department of Internal Medicine, University of Cincinnati , Cincinnati, OH, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Kellie J","family":"Archer","sequence":"additional","affiliation":[{"name":"Division of Biostatistics, College of Public Health, The Ohio State University , Columbus, OH, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"286","published-online":{"date-parts":[[2022,9,30]]},"reference":[{"key":"2022112111110837600_ref1","doi-asserted-by":"crossref","DOI":"10.1007\/978-3-319-40618-3","volume-title":"AJCC Cancer Staging Manual","author":"Amin","year":"2017"},{"key":"2022112111110837600_ref2","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1111\/j.2517-6161.1984.tb01270.x","article-title":"Regression and ordered categorical variables","volume":"46","author":"Anderson","year":"1984","journal-title":"J R Stat Soc Series B Stat Methodol"},{"key":"2022112111110837600_ref3","doi-asserted-by":"crossref","first-page":"1877","DOI":"10.1016\/j.csda.2005.02.013","article-title":"On the estimation of the stereotype regression model","volume":"50","author":"Kuss","year":"2006","journal-title":"Comput Stat Data Anal"},{"key":"2022112111110837600_ref4","doi-asserted-by":"crossref","first-page":"1357","DOI":"10.1002\/sim.2009","article-title":"Prediction of ordinal outcomes when the association between predictors and outcome differs between outcome levels","volume":"24","author":"Lunt","year":"2005","journal-title":"Stat Med"},{"key":"2022112111110837600_ref5","doi-asserted-by":"crossref","first-page":"453","DOI":"10.1182\/blood-2009-07-235358","article-title":"Diagnosis and management of acute myeloid leukemia in adults: recommendations from an international expert panel, on behalf of the European LeukemiaNet","volume":"115","author":"D\u00f6hner","year":"2010","journal-title":"Blood"},{"key":"2022112111110837600_ref6","doi-asserted-by":"crossref","first-page":"424","DOI":"10.1182\/blood-2016-08-733196","article-title":"Diagnosis and management of AML in adults: 2017 ELN recommendations from an international expert panel","volume":"129","author":"D\u00f6hner","year":"2017","journal-title":"Blood"},{"key":"2022112111110837600_ref7","doi-asserted-by":"crossref","DOI":"10.4137\/CIN.S20806","article-title":"ordinalgmifs: An R package for ordinal regression in high-dimensional data settings","volume":"13","author":"Archer","year":"2014","journal-title":"Cancer Inform"},{"key":"2022112111110837600_ref8","doi-asserted-by":"crossref","first-page":"1453","DOI":"10.1002\/sim.8851","article-title":"Bayesian penalized cumulative logit model for high-dimensional data with an ordinal response","volume":"40","author":"Zhang","year":"2021","journal-title":"Stat Med"},{"key":"2022112111110837600_ref9","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1186\/s12859-021-04432-w","article-title":"Bayesian variable selection for high-dimensional data with an ordinal response: identifying genes associated with prognostic risk group in acute myeloid leukemia","volume":"22","author":"Zhang","year":"2021","journal-title":"BMC Bioinformatics"},{"key":"2022112111110837600_ref10","doi-asserted-by":"crossref","first-page":"1665","DOI":"10.1002\/sim.4780131607","article-title":"Alternative models for ordinal logistic regression","volume":"13","author":"Greenland","year":"1994","journal-title":"Stat Med"},{"key":"2022112111110837600_ref11","first-page":"531","article-title":"Transformation of the independent variables","volume":"4","author":"Box","year":"1962","journal-title":"Dent Tech"},{"volume-title":"SAS\/ETS Sser\u2019s Guide","year":"2002","author":"SAS Institute Inc","key":"2022112111110837600_ref12"},{"key":"2022112111110837600_ref13","doi-asserted-by":"crossref","first-page":"3139","DOI":"10.1002\/sim.3693","article-title":"Bayesian inference for the stereotype regression model: application to a case\u2013control study of prostate cancer","volume":"28","author":"Ahn","year":"2009","journal-title":"Stat Med"},{"key":"2022112111110837600_ref14","first-page":"337","article-title":"Adaptive rejection sampling for Gibbs sampling","volume":"41","author":"Gilks","year":"1992","journal-title":"J R Stat Soc Ser C Appl Stat"},{"key":"2022112111110837600_ref15","doi-asserted-by":"crossref","first-page":"457","DOI":"10.1214\/ss\/1177011136","article-title":"Inference from iterative simulation using multiple sequences","volume":"7","author":"Gelman","year":"1992","journal-title":"Stat Sci"},{"key":"2022112111110837600_ref16","doi-asserted-by":"crossref","first-page":"267","DOI":"10.1111\/j.2517-6161.1996.tb02080.x","article-title":"Regression shrinkage and selection via the Lasso","volume":"58","author":"Tibshirani","year":"1996","journal-title":"J R Stat Soc Series B Stat Methodol."},{"key":"2022112111110837600_ref17","doi-asserted-by":"crossref","first-page":"681","DOI":"10.1198\/016214508000000337","article-title":"The Bayesian Lasso","volume":"103","author":"Park","year":"2008","journal-title":"J Am Stat Assoc"},{"key":"2022112111110837600_ref18","doi-asserted-by":"crossref","first-page":"571","DOI":"10.4310\/SII.2014.v7.n4.a12","article-title":"A new Bayesian Lasso","volume":"7","author":"Mallick","year":"2014","journal-title":"Stat Interface"},{"key":"2022112111110837600_ref19","doi-asserted-by":"crossref","first-page":"835","DOI":"10.1093\/biomet\/asp047","article-title":"Bayesian Lasso regression","volume":"96","author":"Hans","year":"2009","journal-title":"Biometrika"},{"key":"2022112111110837600_ref20","doi-asserted-by":"crossref","first-page":"1383","DOI":"10.1198\/jasa.2011.tm09241","article-title":"Elastic net regression modeling with the orthant normal prior","volume":"106","author":"Hans","year":"2011","journal-title":"J Am Stat Assoc"},{"key":"2022112111110837600_ref21","doi-asserted-by":"crossref","first-page":"99","DOI":"10.1111\/j.2517-6161.1974.tb00989.x","article-title":"Scale mixtures of normal distributions","volume":"36","author":"Andrews","year":"1974","journal-title":"J R Stat Soc Series B Stat Methodol"},{"key":"2022112111110837600_ref22","doi-asserted-by":"crossref","first-page":"361","DOI":"10.1007\/s11222-012-9316-x","article-title":"On Bayesian Lasso variable selection and the specification of the shrinkage parameter","volume":"23","author":"Lykou","year":"2013","journal-title":"Stat Comput"},{"key":"2022112111110837600_ref23","doi-asserted-by":"crossref","first-page":"881","DOI":"10.1080\/01621459.1993.10476353","article-title":"Variable selection via Gibbs sampling","volume":"88","author":"George","year":"1993","journal-title":"J Am Stat Assoc"},{"key":"2022112111110837600_ref24","first-page":"65","volume-title":"Variable Selection for Regression Models","author":"Kuo","year":"1998"},{"key":"2022112111110837600_ref25","doi-asserted-by":"crossref","first-page":"27","DOI":"10.1023\/A:1013164120801","article-title":"On Bayesian model and variable selection using MCMC","volume":"12","author":"Dellaportas","year":"2002","journal-title":"Stat Comput."},{"key":"2022112111110837600_ref26","doi-asserted-by":"crossref","first-page":"203","DOI":"10.1007\/s11222-009-9158-3","article-title":"Bayesian regularisation in structured additive regression: a unifying perspective on shrinkage, smoothing and predictor selection","volume":"20","author":"Fahrmeir","year":"2010","journal-title":"Stat Comput"},{"key":"2022112111110837600_ref27","doi-asserted-by":"crossref","first-page":"870","DOI":"10.1214\/009053604000000238","article-title":"Optimal predictive model selection","volume":"32","author":"Barbieri","year":"2004","journal-title":"Ann Stat"},{"key":"2022112111110837600_ref28","doi-asserted-by":"crossref","first-page":"221","DOI":"10.1007\/s11222-009-9160-9","article-title":"Model uncertainty and variable selection in Bayesian Lasso regression","volume":"20","author":"Hans","year":"2010","journal-title":"Stat Comput."},{"key":"2022112111110837600_ref29","doi-asserted-by":"crossref","first-page":"587","DOI":"10.1111\/j.1541-0420.2011.01680.x","article-title":"Logistic Bayesian LASSO for identifying association with rare haplotypes and application to age-related macular degeneration","volume":"68","author":"Biswas","year":"2012","journal-title":"Biometrics"},{"key":"2022112111110837600_ref30","doi-asserted-by":"crossref","first-page":"773","DOI":"10.1080\/01621459.1995.10476572","article-title":"Bayes factors","volume":"90","author":"Kass","year":"1995","journal-title":"J Am Stat Assoc"},{"volume-title":"Language and Environment for Stat Comput","year":"2019","author":"R Core Team. R: A","key":"2022112111110837600_ref31"},{"volume-title":"rjags: Bayesian Graphical Models using MCMC","year":"2019","author":"Plummer","key":"2022112111110837600_ref32"},{"volume-title":"Ohio Supercomputer Center","year":"1987","author":"Ohio Supercomputer Center","key":"2022112111110837600_ref33"},{"key":"2022112111110837600_ref34","doi-asserted-by":"crossref","first-page":"3088","DOI":"10.1182\/blood-2008-09-179895","article-title":"Double CEBPA mutations, but not single CEBPA mutations, define a subgroup of acute myeloid leukemia with a distinctive gene expression profile that is uniquely associated with a favorable outcome","volume":"113","author":"Wouters","year":"2009","journal-title":"Blood"},{"key":"2022112111110837600_ref35","doi-asserted-by":"crossref","first-page":"2469","DOI":"10.1182\/blood-2010-09-307280","article-title":"Prognostic impact, concurrent genetic mutations, and gene expression features of AML with CEBPA mutations in a cohort of 1182 cytogenetically normal AML patients: further evidence for CEBPA double mutant AML as a distinctive disease entity","volume":"117","author":"Taskesen","year":"2011","journal-title":"Blood"},{"issue":"Suppl 4","key":"2022112111110837600_ref36","doi-asserted-by":"crossref","first-page":"S5","DOI":"10.1186\/1471-2105-16-S4-S5","article-title":"Integration of gene expression and DNA-methylation profiles improves molecular subtype classification in acute myeloid leukemia","volume":"16","author":"Taskesen","year":"2015","journal-title":"BMC Bioinformatics"},{"key":"2022112111110837600_ref37","doi-asserted-by":"crossref","first-page":"1846","DOI":"10.1093\/bioinformatics\/btm254","article-title":"GEOquery: a bridge between the Gene Expression Omnibus (GEO) and BioConductor","volume":"23","author":"Davis","year":"2007","journal-title":"Bioinformatics"},{"issue":"3","key":"2022112111110837600_ref38","doi-asserted-by":"crossref","first-page":"307","DOI":"10.1093\/bioinformatics\/btg405","article-title":"affy\u2014analysis of Affymetrix GeneChip data at the probe level","volume":"20","author":"Gautier","year":"2004","journal-title":"Bioinformatics"},{"key":"2022112111110837600_ref39","doi-asserted-by":"crossref","first-page":"249","DOI":"10.1093\/biostatistics\/4.2.249","article-title":"Exploration, normalization, and summaries of high density oligonucleotide array probe level data","volume":"4","author":"Irizarry","year":"2003","journal-title":"Biostatistics"},{"key":"2022112111110837600_ref40","doi-asserted-by":"crossref","first-page":"185","DOI":"10.1093\/bioinformatics\/19.2.185","article-title":"A comparison of normalization methods for high density oligonucleotide array data based on variance and bias","volume":"19","author":"Bolstad","year":"2003","journal-title":"Bioinformatics"},{"key":"2022112111110837600_ref41","doi-asserted-by":"crossref","first-page":"e15","DOI":"10.1093\/nar\/gng015","article-title":"Summaries of Affymetrix GeneChip probe level data","volume":"31","author":"Irizarry","year":"2003","journal-title":"Nucleic Acids Res"},{"volume-title":"Proc. SPIE 4266, Microarrays: Optical Technologies and Informatics","year":"4 2001","author":"Wm","key":"2022112111110837600_ref42"},{"key":"2022112111110837600_ref43","doi-asserted-by":"crossref","first-page":"1593","DOI":"10.1093\/bioinformatics\/18.12.1593","article-title":"Analysis of high density expression microarrays with signed-rank call algorithms","volume":"18","author":"Wm","year":"2002","journal-title":"Bioinformatics"},{"journal-title":"Statistical algorithms description document","year":"2002","author":"Affymetrix","key":"2022112111110837600_ref44"},{"volume-title":"caret: Classification and Regression Training","year":"2020","author":"Kuhn","key":"2022112111110837600_ref45"},{"key":"2022112111110837600_ref46","doi-asserted-by":"crossref","first-page":"10","DOI":"10.14806\/ej.17.1.200","article-title":"Cutadapt removes adapter sequences from high-throughput sequencing reads","volume":"17","author":"Martin","year":"2011","journal-title":"EMBNet J"},{"key":"2022112111110837600_ref47","doi-asserted-by":"crossref","first-page":"15","DOI":"10.1093\/bioinformatics\/bts635","article-title":"STAR: ultrafast universal RNA-seq aligner","volume":"29","author":"Dobin","year":"2013","journal-title":"Bioinformatics"},{"key":"2022112111110837600_ref48","doi-asserted-by":"crossref","first-page":"166","DOI":"10.1093\/bioinformatics\/btu638","article-title":"HTSeq-a Python framework to work with high-throughput sequencing data","volume":"31","author":"Anders","year":"2015","journal-title":"Bioinformatics"},{"key":"2022112111110837600_ref49","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1186\/gb-2014-15-2-r29","article-title":"voom: Precision weights unlock linear model analysis tools for RNA-seq read counts","volume":"15","author":"Law","year":"2014","journal-title":"Genome Biol"},{"key":"2022112111110837600_ref50","doi-asserted-by":"crossref","first-page":"1","DOI":"10.18637\/jss.v032.i10","article-title":"The VGAM package for categorical data analysis","volume":"32","author":"Yee","year":"2010","journal-title":"J Stat Softw"},{"key":"2022112111110837600_ref51","doi-asserted-by":"crossref","first-page":"526","DOI":"10.1038\/s41586-018-0623-z","article-title":"Functional genomic landscape of acute myeloid leukaemia","volume":"562","author":"Tyner","year":"2018","journal-title":"Nature"},{"key":"2022112111110837600_ref52","doi-asserted-by":"crossref","DOI":"10.1093\/nar\/gkv1507","article-title":"TCGAbiolinks: an R\/Bioconductor package for integrative analysis of TCGA data","author":"Colaprico","year":"2016","journal-title":"Nucleic Acids Res"},{"key":"2022112111110837600_ref53","doi-asserted-by":"crossref","first-page":"1542","DOI":"10.12688\/f1000research.8923.1","article-title":"Analyze cancer genomics and epigenomics data using Bioconductor packages","volume":"5","author":"Silva","year":"2016","journal-title":"F1000Res"},{"issue":"3","key":"2022112111110837600_ref54","doi-asserted-by":"crossref","DOI":"10.1371\/journal.pcbi.1006701","article-title":"New functionalities in the TCGAbiolinks package for the study and integration of cancer data from GDC and GTEx","volume":"15","author":"Mounir","year":"2019","journal-title":"PLoS Comput Biol"},{"key":"2022112111110837600_ref55","doi-asserted-by":"crossref","first-page":"465","DOI":"10.1093\/biomet\/asq017","article-title":"The horseshoe estimator for sparse signals","volume":"97","author":"Carvalho","year":"2010","journal-title":"Biometrika"},{"issue":"5","key":"2022112111110837600_ref56","doi-asserted-by":"crossref","first-page":"971","DOI":"10.1109\/TCBB.2015.2478454","article-title":"Supervised, unsupervised, and semi-supervised feature selection: a review on gene selection","volume":"13","author":"Ang","year":"2015","journal-title":"IEEE\/ACM Trans Comput Biol Bioinform"},{"issue":"6","key":"2022112111110837600_ref57","doi-asserted-by":"crossref","DOI":"10.1093\/bib\/bbab295","article-title":"Selecting gene features for unsupervised analysis of single-cell gene expression data","volume":"22","author":"Sheng","year":"2021","journal-title":"Brief Bioinform"}],"container-title":["Briefings in Bioinformatics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/academic.oup.com\/bib\/article-pdf\/23\/6\/bbac414\/47143899\/bbac414.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/academic.oup.com\/bib\/article-pdf\/23\/6\/bbac414\/47143899\/bbac414.pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,10,2]],"date-time":"2024-10-02T19:50:55Z","timestamp":1727898655000},"score":1,"resource":{"primary":{"URL":"https:\/\/academic.oup.com\/bib\/article\/doi\/10.1093\/bib\/bbac414\/6731715"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,9,30]]},"references-count":57,"journal-issue":{"issue":"6","published-print":{"date-parts":[[2022,11,19]]}},"URL":"https:\/\/doi.org\/10.1093\/bib\/bbac414","relation":{},"ISSN":["1467-5463","1477-4054"],"issn-type":[{"type":"print","value":"1467-5463"},{"type":"electronic","value":"1477-4054"}],"subject":[],"published-other":{"date-parts":[[2022,11]]},"published":{"date-parts":[[2022,9,30]]},"article-number":"bbac414"}}