{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,8,2]],"date-time":"2026-08-02T10:22:53Z","timestamp":1785666173226,"version":"3.56.0"},"reference-count":13,"publisher":"Oxford University Press (OUP)","issue":"7","license":[{"start":{"date-parts":[[2018,9,5]],"date-time":"2018-09-05T00:00:00Z","timestamp":1536105600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/academic.oup.com\/journals\/pages\/open_access\/funder_policies\/chorus\/standard_publication_model"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2019,4,1]]},"abstract":"<jats:title>Abstract<\/jats:title><jats:sec><jats:title>Summary<\/jats:title><jats:p>VarSelLCM allows a full model selection (detection of the relevant features for clustering and selection of the number of clusters) in model-based clustering, according to classical information criteria. Data to be analyzed can be composed of continuous, integer and\/or categorical features. Moreover, missing values are managed, without any pre-processing, by the model used to cluster with the assumption that values are missing completely at random. Thus, VarSelLCM also allows data imputation by using mixture models. A Shiny application is implemented to easily interpret the clustering results.<\/jats:p><\/jats:sec><jats:sec><jats:title>Availability and implementation<\/jats:title><jats:p>VarSelLCM is available to download at https:\/\/CRAN.R-project.org\/package=VarSelLCM\/.<\/jats:p><\/jats:sec><jats:sec><jats:title>Tutorial<\/jats:title><jats:p>vignette is available online at http:\/\/varsellcm.r-forge.r-project.org\/<\/jats:p><\/jats:sec><jats:sec><jats:title>Supplementary information<\/jats:title><jats:p>Supplementary data are available at Bioinformatics online.<\/jats:p><\/jats:sec>","DOI":"10.1093\/bioinformatics\/bty786","type":"journal-article","created":{"date-parts":[[2018,9,4]],"date-time":"2018-09-04T12:21:56Z","timestamp":1536063716000},"page":"1255-1257","source":"Crossref","is-referenced-by-count":61,"title":["VarSelLCM: an R\/C++ package for variable selection in model-based clustering of mixed-data with missing values"],"prefix":"10.1093","volume":"35","author":[{"given":"Matthieu","family":"Marbac","sequence":"first","affiliation":[{"name":"CREST, Ensai, Bruz, France"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Mohammed","family":"Sedki","sequence":"additional","affiliation":[{"name":"University of Paris-Sud and UMR Inserm-1181, Paris, France"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"286","published-online":{"date-parts":[[2018,9,5]]},"reference":[{"key":"2023013107271945900_bty786-B1","doi-asserted-by":"crossref","first-page":"719","DOI":"10.1109\/34.865189","article-title":"Assessing a mixture model for clustering with the integrated completed likelihood","volume":"22","author":"Biernacki","year":"2000","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell"},{"key":"2023013107271945900_bty786-B2","doi-asserted-by":"crossref","first-page":"561","DOI":"10.1016\/S0167-9473(02)00163-9","article-title":"Choosing starting values for the EM algorithm for getting the highest likelihood in multivariate Gaussian mixture models","volume":"41","author":"Biernacki","year":"2003","journal-title":"Comput. Stat. Data Anal"},{"key":"2023013107271945900_bty786-B3","author":"Chang","year":"2017"},{"key":"2023013107271945900_bty786-B4","author":"Fop","year":"2017"},{"key":"2023013107271945900_bty786-B5","first-page":"18","volume-title":"Statist. Surv","author":"Fop","year":"2017"},{"key":"2023013107271945900_bty786-B6","doi-asserted-by":"crossref","first-page":"443","DOI":"10.1111\/j.2517-6161.1990.tb01798.x","article-title":"On use of the EM for penalized likelihood estimation","volume":"52","author":"Green","year":"1990","journal-title":"J. Royal Stat. Soc"},{"key":"2023013107271945900_bty786-B7","volume-title":"Statistical Analysis with Missing Data","author":"Little","year":"2014"},{"key":"2023013107271945900_bty786-B8","doi-asserted-by":"crossref","first-page":"1049","DOI":"10.1007\/s11222-016-9670-1","article-title":"Variable selection for model-based clustering using the integrated complete-data likelihood","volume":"27","author":"Marbac","year":"2017","journal-title":"Stat. Comput"},{"key":"2023013107271945900_bty786-B9","author":"Marbac","year":"2018"},{"key":"2023013107271945900_bty786-B10","doi-asserted-by":"crossref","first-page":"168","DOI":"10.1198\/016214506000000113","article-title":"Variable selection for model-based clustering","volume":"101","author":"Raftery","year":"2006","journal-title":"J. Am. Stat. Assoc"},{"key":"2023013107271945900_bty786-B11","doi-asserted-by":"crossref","first-page":"461","DOI":"10.1214\/aos\/1176344136","article-title":"Estimating the dimension of a model","volume":"6","author":"Schwarz","year":"1978","journal-title":"Ann. Stat"},{"key":"2023013107271945900_bty786-B12","doi-asserted-by":"crossref","first-page":"1","DOI":"10.18637\/jss.v084.i01","article-title":"clustvarsel: a package implementing variable selection for model-based clustering in R","volume":"84","author":"Scrucca","year":"2018","journal-title":"J. Stat. Softw"},{"key":"2023013107271945900_bty786-B13","author":"Witten","year":"2013"}],"container-title":["Bioinformatics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/academic.oup.com\/bioinformatics\/article-pdf\/35\/7\/1255\/48967518\/bioinformatics_35_7_1255.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/academic.oup.com\/bioinformatics\/article-pdf\/35\/7\/1255\/48967518\/bioinformatics_35_7_1255.pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,7,9]],"date-time":"2024-07-09T22:44:40Z","timestamp":1720565080000},"score":1,"resource":{"primary":{"URL":"https:\/\/academic.oup.com\/bioinformatics\/article\/35\/7\/1255\/5091183"}},"subtitle":[],"editor":[{"given":"Jonathan","family":"Wren","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"editor"}]}],"short-title":[],"issued":{"date-parts":[[2018,9,5]]},"references-count":13,"journal-issue":{"issue":"7","published-print":{"date-parts":[[2019,4,1]]}},"URL":"https:\/\/doi.org\/10.1093\/bioinformatics\/bty786","relation":{},"ISSN":["1367-4803","1367-4811"],"issn-type":[{"value":"1367-4803","type":"print"},{"value":"1367-4811","type":"electronic"}],"subject":[],"published-other":{"date-parts":[[2019,4,1]]},"published":{"date-parts":[[2018,9,5]]}}}