{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,1,22]],"date-time":"2026-01-22T08:25:07Z","timestamp":1769070307513,"version":"3.49.0"},"reference-count":34,"publisher":"Springer Science and Business Media LLC","issue":"1","content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["BMC Bioinformatics"],"published-print":{"date-parts":[[2014,12]]},"abstract":"<jats:title>Abstract<\/jats:title>\n          <jats:sec>\n            <jats:title>Background<\/jats:title>\n            <jats:p>High-throughput sequencing allows the detection and quantification of frequencies of somatic single nucleotide variants (SNV) in heterogeneous tumor cell populations. In some cases, the evolutionary history and population frequency of the subclonal lineages of tumor cells present in the sample can be reconstructed from these SNV frequency measurements. But automated methods to do this reconstruction are not available and the conditions under which reconstruction is possible have not been described.<\/jats:p>\n          <\/jats:sec>\n          <jats:sec>\n            <jats:title>Results<\/jats:title>\n            <jats:p>We describe the conditions under which the evolutionary history can be uniquely reconstructed from SNV frequencies from single or multiple samples from the tumor population and we introduce a new statistical model, <jats:italic>PhyloSub<\/jats:italic>, that infers the phylogeny and genotype of the major subclonal lineages represented in the population of cancer cells. It uses a Bayesian nonparametric prior over trees that groups SNVs into major subclonal lineages and automatically estimates the number of lineages and their ancestry. We sample from the joint posterior distribution over trees to identify evolutionary histories and cell population frequencies that have the highest probability of generating the observed SNV frequency data. When multiple phylogenies are consistent with a given set of SNV frequencies, PhyloSub represents the uncertainty in the tumor phylogeny using a \u201cpartial order plot\u201d. Experiments on a simulated dataset and two real datasets comprising tumor samples from acute myeloid leukemia and chronic lymphocytic leukemia patients demonstrate that PhyloSub can infer both linear (or chain) and branching lineages and its inferences are in good agreement with ground truth, where it is available.<\/jats:p>\n          <\/jats:sec>\n          <jats:sec>\n            <jats:title>Conclusions<\/jats:title>\n            <jats:p>PhyloSub can be applied to frequencies of any \u201cbinary\u201d somatic mutation, including SNVs as well as small insertions and deletions. The PhyloSub and partial order plot software is available from <jats:ext-link xmlns:xlink=\"http:\/\/www.w3.org\/1999\/xlink\" xlink:href=\"https:\/\/github.com\/morrislab\/phylosub\/\" ext-link-type=\"uri\">https:\/\/github.com\/morrislab\/phylosub\/<\/jats:ext-link>.<\/jats:p>\n          <\/jats:sec>","DOI":"10.1186\/1471-2105-15-35","type":"journal-article","created":{"date-parts":[[2014,2,1]],"date-time":"2014-02-01T12:01:04Z","timestamp":1391256064000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":209,"title":["Inferring clonal evolution of tumors from single nucleotide somatic mutations"],"prefix":"10.1186","volume":"15","author":[{"given":"Wei","family":"Jiao","sequence":"first","affiliation":[]},{"given":"Shankar","family":"Vembu","sequence":"additional","affiliation":[]},{"given":"Amit G","family":"Deshwar","sequence":"additional","affiliation":[]},{"given":"Lincoln","family":"Stein","sequence":"additional","affiliation":[]},{"given":"Quaid","family":"Morris","sequence":"additional","affiliation":[]}],"member":"297","published-online":{"date-parts":[[2014,2,1]]},"reference":[{"key":"6294_CR1","doi-asserted-by":"publisher","first-page":"57","DOI":"10.1016\/S0092-8674(00)81683-9","volume":"100","author":"D Hanahan","year":"2000","unstructured":"Hanahan D, Weinberg RA: The hallmarks of cancer. Cell. 2000, 100: 57-70. 10.1016\/S0092-8674(00)81683-9.","journal-title":"Cell"},{"issue":"5","key":"6294_CR2","doi-asserted-by":"publisher","first-page":"646","DOI":"10.1016\/j.cell.2011.02.013","volume":"144","author":"D Hanahan","year":"2011","unstructured":"Hanahan D, Weinberg RA: Hallmarks of cancer: The next generation. Cell. 2011, 144 (5): 646-674. 10.1016\/j.cell.2011.02.013.","journal-title":"Cell"},{"issue":"127","key":"6294_CR3","first-page":"127ps10","volume":"4","author":"TA Yap","year":"2012","unstructured":"Yap TA, Gerlinger M, Futreal PA, Pusztai L, Swanton C: Intratumor heterogeneity: Seeing the wood for the trees. Sci Transl Med. 2012, 4 (127): 127ps10-","journal-title":"Sci Transl Med"},{"issue":"7330","key":"6294_CR4","doi-asserted-by":"publisher","first-page":"314","DOI":"10.1038\/nature09781","volume":"469","author":"JE Visvader","year":"2011","unstructured":"Visvader JE: Cells of origin in cancer. Nature. 2011, 469 (7330): 314-322. 10.1038\/nature09781.","journal-title":"Nature"},{"issue":"3","key":"6294_CR5","doi-asserted-by":"publisher","first-page":"267","DOI":"10.1016\/j.molonc.2010.04.010","volume":"4","author":"NE Navin","year":"2010","unstructured":"Navin NE, Hicks J: Tracing the tumor lineage. Mol Oncol. 2010, 4 (3): 267-283. 10.1016\/j.molonc.2010.04.010.","journal-title":"Mol Oncol"},{"issue":"5906","key":"6294_CR6","doi-asserted-by":"publisher","first-page":"1377","DOI":"10.1126\/science.1164266","volume":"322","author":"CG Mullighan","year":"2008","unstructured":"Mullighan CG, Phillips LA, Su X, Ma J, Miller CB, Shurtleff SA, Downing JR: Genomic analysis of the clonal origins of relapsed acute lymphoblastic leukemia. Science. 2008, 322 (5906): 1377-1380. 10.1126\/science.1164266.","journal-title":"Science"},{"issue":"10","key":"6294_CR7","doi-asserted-by":"publisher","first-page":"883","DOI":"10.1056\/NEJMoa1113205","volume":"366","author":"M Gerlinger","year":"2012","unstructured":"Gerlinger M, Rowan AJ, Horswell S, Larkin J, Endesfelder D, Gronroos E, Martinez P, Matthews N, Stewart A, Tarpey P, Varela I, Phillimore B, Begum S, McDonald NQ, Butler A, Jones D, Raine K, Latimer C, Santos CR, Nohadani M, Eklund AC, Spencer-Dene B, Clark G, Pickering L, Stamp G, Gore M, Szallasi Z, Downward J, Futreal PA, Swanton C: Intratumor heterogeneity and branched evolution revealed by multiregion sequencing. N Engl J Med. 2012, 366 (10): 883-892. 10.1056\/NEJMoa1113205.","journal-title":"N Engl J Med"},{"key":"6294_CR8","first-page":"105","volume":"1805","author":"A Marusyk","year":"2010","unstructured":"Marusyk A, Polyak K: Tumor heterogeneity: Causes and consequences. Biochim Biophys Acta. 2010, 1805: 105-117.","journal-title":"Biochim Biophys Acta"},{"issue":"20","key":"6294_CR9","doi-asserted-by":"publisher","first-page":"4191","DOI":"10.1182\/blood-2012-05-433540","volume":"120","author":"A Schuh","year":"2012","unstructured":"Schuh A, Becq J, Humphray S, Alexa A, Burns A, Clifford R, Feller SM, Grocock R, Henderson S, Khrebtukova I, Kingsbury Z, Luo S, McBride D, Murray L, Menju T, Timbs A, Ross M, Taylor J, Bentley D: Monitoring chronic lymphocytic leukemia progression by whole genome sequencing reveals heterogeneous clonal evolution patterns. Blood. 2012, 120 (20): 4191-4196. 10.1182\/blood-2012-05-433540.","journal-title":"Blood"},{"issue":"7403","key":"6294_CR10","first-page":"617","volume":"486","author":"SP Shah","year":"2012","unstructured":"Shah SP, Roth A, Goya R, Oloumi A, Ha G, Zhao Y, Turashvili G, Ding J, Tse K, Haffari G, Bashashati A, Prentice LM, Khattra J, Burleigh A, Yap D, Bernard V, McPherson A, Shumansky K, Crisan A, Giuliany R, Heravi-Moussavi A, Rosner J, Lai D, Birol I, Varhol R, Tam A, Dhalla N, Zeng T, Ma K, Chan SK, et al: The clonal and mutational evolution spectrum of primary triple-negative breast cancers. Nature. 2012, 486 (7403): 617-656.","journal-title":"Nature"},{"issue":"5","key":"6294_CR11","doi-asserted-by":"publisher","first-page":"413","DOI":"10.1038\/nbt.2203","volume":"30","author":"SL Carter","year":"2012","unstructured":"Carter SL, Cibulskis K, Helman E, McKenna A, Shen H, Zack T, Laird PW, Onofrio RC, Winckler W, Weir BA, Beroukhim R, Pellman D, Levine DA, Lander ES, Meyerson M, Getz G: Absolute quantification of somatic DNA alterations in human cancer. Nat Biotechnol. 2012, 30 (5): 413-421. 10.1038\/nbt.2203.","journal-title":"Nat Biotechnol"},{"issue":"4","key":"6294_CR12","doi-asserted-by":"publisher","first-page":"714","DOI":"10.1016\/j.cell.2013.01.019","volume":"152","author":"DA Landau","year":"2013","unstructured":"Landau DA, Carter SL, Stojanov P, McKenna A, Stevenson K, Lawrence MS, Sougnez C, Stewart C, Sivachenko A, Wang L, Wan Y, Zhang W, Shukla SA, Vartanov A, Fernandes SM, Saksena G, Cibulskis K, Tesar B, Gabriel S, Hacohen N, Meyerson M, Lander ES, Neuberg D, Brown JR, Getz G, Wu CJ: Evolution and impact of subclonal mutations in chronic lymphocytic leukemia. Cell. 2013, 152 (4): 714-726. 10.1016\/j.cell.2013.01.019.","journal-title":"Cell"},{"issue":"7","key":"6294_CR13","doi-asserted-by":"publisher","first-page":"R80","DOI":"10.1186\/gb-2013-14-7-r80","volume":"14","author":"L Oesper","year":"2013","unstructured":"Oesper L, Mahmoody A, Raphael BJ: THetA: Inferring intra-tumor heterogeneity from high-throughput DNA sequencing data. Genome Biol. 2013, 14 (7): R80-10.1186\/gb-2013-14-7-r80.","journal-title":"Genome Biol"},{"issue":"5","key":"6294_CR14","doi-asserted-by":"publisher","first-page":"323","DOI":"10.1038\/nrc3261","volume":"12","author":"A Marusyk","year":"2012","unstructured":"Marusyk A, Almendro V, Polyak K: Intra-tumour heterogeneity: A looking glass for cancer?. Nat Rev Cancer. 2012, 12 (5): 323-334. 10.1038\/nrc3261.","journal-title":"Nat Rev Cancer"},{"issue":"10","key":"6294_CR15","doi-asserted-by":"publisher","first-page":"685","DOI":"10.1038\/nrg2841","volume":"11","author":"M Meyerson","year":"2010","unstructured":"Meyerson M, Gabriel S, Getz G: Advances in understanding cancer genomes through second-generation sequencing. Nat Rev Genet. 2010, 11 (10): 685-696. 10.1038\/nrg2841.","journal-title":"Nat Rev Genet"},{"key":"6294_CR16","doi-asserted-by":"publisher","first-page":"994","DOI":"10.1016\/j.cell.2012.04.023","volume":"149","author":"S Nik-Zainal","year":"2012","unstructured":"Nik-Zainal S, Loo PV, Wedge DC, Alexandrov LB, Greenman CD, Lau KW, Raine K, Jones D, Marshall J, Ramakrishna M, Shlien A, Cooke SL, Hinton J, Menzies A, Stebbings LA, Leroy C, Jia M, Rance R, Mudie LJ, Gamble SJ, Stephens PJ, McLaren S, Tarpey PS, Papaemmanuil E, Davies HR, Varela I, McBride DJ, Bignell GR, Leung K, Butler AP, et al: The life history of 21 breast cancers. Cell. 2012, 149: 994-1007. 10.1016\/j.cell.2012.04.023.","journal-title":"Cell"},{"issue":"149","key":"6294_CR17","first-page":"149ra118","volume":"4","author":"M Jan","year":"2012","unstructured":"Jan M, Snyder TM, Corces-Zimmerman MR, Vyas P, Weissman IL, Quake SR, Majeti R: Clonal evolution of preleukemic hematopoietic stem cells precedes human acute myeloid leukemia. Sci Transl Med. 2012, 4 (149): 149ra118-","journal-title":"Sci Transl Med"},{"issue":"35","key":"6294_CR18","doi-asserted-by":"publisher","first-page":"13081","DOI":"10.1073\/pnas.0801523105","volume":"105","author":"PJ Campbell","year":"2008","unstructured":"Campbell PJ, Pleasance ED, Stephens PJ, Dicks E, Rance R, Goodhead I, Follows GA, Green AR, Futreal PA, Stratton MR: Subclonal phylogenetic structures in cancer revealed by ultra-deep sequencing. Proc Nat Acad Sci. 2008, 105 (35): 13081-13086. 10.1073\/pnas.0801523105.","journal-title":"Proc Nat Acad Sci"},{"key":"6294_CR19","volume-title":"Proceedings of the 24th Annual Conference on Neural Information Processing Systems","author":"RP Adams","year":"2010","unstructured":"Adams RP, Ghahramani Z, Jordan MI: Tree-structured stick breaking for hierarchical data. Proceedings of the 24th Annual Conference on Neural Information Processing Systems. 2010,"},{"issue":"2","key":"6294_CR20","doi-asserted-by":"publisher","first-page":"237","DOI":"10.1016\/j.semcdb.2011.12.008","volume":"23","author":"J Brosnan","year":"2012","unstructured":"Brosnan J, Iacobuzio-Donahue C: A new branch on the tree: Next-generation sequencing in the study of cancer evolution. Semin Cell Dev Biol. 2012, 23 (2): 237-242. 10.1016\/j.semcdb.2011.12.008.","journal-title":"Semin Cell Dev Biol"},{"issue":"4","key":"6294_CR21","doi-asserted-by":"crossref","first-page":"893","DOI":"10.1093\/genetics\/61.4.893","volume":"61","author":"M Kimura","year":"1969","unstructured":"Kimura M: The number of heterozygous nucleotide sites maintained in a finite population due to steady flux of mutations. Genetics. 1969, 61 (4): 893-","journal-title":"Genetics"},{"issue":"2","key":"6294_CR22","doi-asserted-by":"publisher","first-page":"183","DOI":"10.1016\/0040-5809(83)90013-8","volume":"23","author":"RR Hudson","year":"1983","unstructured":"Hudson RR: Properties of a neutral allele model with intragenic recombination. Theor Popul Biol. 1983, 23 (2): 183-201. 10.1016\/0040-5809(83)90013-8.","journal-title":"Theor Popul Biol"},{"issue":"23","key":"6294_CR23","doi-asserted-by":"publisher","first-page":"2406","DOI":"10.1056\/NEJMoa044190","volume":"352","author":"TM Shattuck","year":"2005","unstructured":"Shattuck TM, Westra WH, Ladenson PW, Arnold A: Independent clonal origins of distinct tumor foci in multifocal papillary thyroid carcinoma. N Engl J Med. 2005, 352 (23): 2406-2412. 10.1056\/NEJMoa044190.","journal-title":"N Engl J Med"},{"key":"6294_CR24","unstructured":"Graphviz - Graph visualization software. http:\/\/graphviz.org,"},{"key":"6294_CR25","doi-asserted-by":"publisher","DOI":"10.1090\/mbk\/058","volume-title":"Markov chains and mixing times","author":"DA Levin","year":"2008","unstructured":"Levin DA, Peres Y, Wilmer EL: Markov chains and mixing times. 2008, AMS Bookstore"},{"issue":"3","key":"6294_CR26","doi-asserted-by":"publisher","first-page":"R25","DOI":"10.1186\/gb-2010-11-3-r25","volume":"11","author":"MD Robinson","year":"2010","unstructured":"Robinson MD, Oshlack A: A scaling normalization method for differential expression analysis of RNA-seq data. Genome Biol. 2010, 11 (3): R25-10.1186\/gb-2010-11-3-r25.","journal-title":"Genome Biol"},{"key":"6294_CR27","volume-title":"Encyclopedia of Machine Learning","author":"YW Teh","year":"2010","unstructured":"Teh YW: Dirichlet processes. Encyclopedia of Machine Learning. 2010, Springer,"},{"issue":"6","key":"6294_CR28","doi-asserted-by":"publisher","first-page":"1152","DOI":"10.1214\/aos\/1176342871","volume":"2","author":"CE Antoniak","year":"1974","unstructured":"Antoniak CE: Mixtures of Dirichlet processes with applications to Bayesian nonparametric problems. Ann Stat. 1974, 2 (6): 1152-1174. 10.1214\/aos\/1176342871.","journal-title":"Ann Stat"},{"key":"6294_CR29","first-page":"639","volume":"4","author":"J Sethuraman","year":"1994","unstructured":"Sethuraman J: A constructive definition of Dirichlet process. Stat Sinica. 1994, 4: 639-650.","journal-title":"Stat Sinica"},{"key":"6294_CR30","unstructured":"PyClone. http:\/\/compbio.bccrc.ca\/software\/pyclone\/,"},{"key":"6294_CR31","unstructured":"Minka TP: Estimating a Dirichlet distribution. Tech. rep., Microsoft Research Cambridge, 2012, http:\/\/citeseerx.ist.psu.edu\/viewdoc\/summary?doi=10.1.1.220.175,"},{"issue":"1\u20133","key":"6294_CR32","doi-asserted-by":"publisher","first-page":"89","DOI":"10.1023\/B:MACH.0000033116.57574.95","volume":"56","author":"N Bansal","year":"2004","unstructured":"Bansal N, Blum A, Chawla S: Correlation clustering. Mach Learn. 2004, 56 (1\u20133): 89-113.","journal-title":"Mach Learn"},{"key":"6294_CR33","doi-asserted-by":"publisher","first-page":"112","DOI":"10.1111\/j.2041-210X.2011.00131.x","volume":"3","author":"WA Link","year":"2012","unstructured":"Link WA, Eaton MJ: On thinning of chains in MCMC. Methods Ecol Evol. 2012, 3: 112-115. 10.1111\/j.2041-210X.2011.00131.x.","journal-title":"Methods Ecol Evol"},{"key":"6294_CR34","first-page":"7","volume":"6","author":"M Plummer","year":"2006","unstructured":"Plummer M, Best N, Cowles K, Vines K: CODA: Convergence diagnosis and output analysis for MCMC. R News. 2006, 6: 7-11.","journal-title":"R News"}],"container-title":["BMC Bioinformatics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1186\/1471-2105-15-35.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2021,9,2]],"date-time":"2021-09-02T02:59:40Z","timestamp":1630551580000},"score":1,"resource":{"primary":{"URL":"https:\/\/bmcbioinformatics.biomedcentral.com\/articles\/10.1186\/1471-2105-15-35"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2014,2,1]]},"references-count":34,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2014,12]]}},"alternative-id":["6294"],"URL":"https:\/\/doi.org\/10.1186\/1471-2105-15-35","relation":{},"ISSN":["1471-2105"],"issn-type":[{"value":"1471-2105","type":"electronic"}],"subject":[],"published":{"date-parts":[[2014,2,1]]},"assertion":[{"value":"21 May 2013","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"24 January 2014","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"1 February 2014","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}}],"article-number":"35"}}