{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,10,17]],"date-time":"2025-10-17T13:41:19Z","timestamp":1760708479947},"reference-count":28,"publisher":"Oxford University Press (OUP)","issue":"1","content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2013,1,1]]},"abstract":"<jats:title>Abstract<\/jats:title>\n               <jats:p>Motivation: Given the current costs of next-generation sequencing, large studies carry out low-coverage sequencing followed by application of methods that leverage linkage disequilibrium to infer genotypes. We propose a novel method that assumes study samples are sequenced at low coverage and genotyped on a genome-wide microarray, as in the 1000 Genomes Project (1KGP). We assume polymorphic sites have been detected from the sequencing data and that genotype likelihoods are available at these sites. We also assume that the microarray genotypes have been phased to construct a haplotype scaffold. We then phase each polymorphic site using an MCMC algorithm that iteratively updates the unobserved alleles based on the genotype likelihoods at that site and local haplotype information. We use a multivariate normal model to capture both allele frequency and linkage disequilibrium information around each site. When sequencing data are available from trios, Mendelian transmission constraints are easily accommodated into the updates. The method is highly parallelizable, as it analyses one position at a time.<\/jats:p>\n               <jats:p>Results: We illustrate the performance of the method compared with other methods using data from Phase 1 of the 1KGP in terms of genotype accuracy, phasing accuracy and downstream imputation performance. We show that the haplotype panel we infer in African samples, which was based on a trio-phased scaffold, increases downstream imputation accuracy for rare variants (R2 increases by &amp;gt;0.05 for minor allele frequency &amp;lt;1%), and this will translate into a boost in power to detect associations. These results highlight the value of incorporating microarray genotypes when calling variants from next-generation sequence data.<\/jats:p>\n               <jats:p>Availability: The method (called MVNcall) is implemented in a C++ program and is available from http:\/\/www.stats.ox.ac.uk\/\u223cmarchini\/#software.<\/jats:p>\n               <jats:p>Contact: \u00a0marchini@stats.ox.ac.uk<\/jats:p>\n               <jats:p>Supplementary information: \u00a0Supplementary data are available at Bioinformatics online.<\/jats:p>","DOI":"10.1093\/bioinformatics\/bts632","type":"journal-article","created":{"date-parts":[[2012,10,24]],"date-time":"2012-10-24T04:55:41Z","timestamp":1351054541000},"page":"84-91","source":"Crossref","is-referenced-by-count":46,"title":["Genotype calling and phasing using next-generation sequencing reads and a haplotype scaffold"],"prefix":"10.1093","volume":"29","author":[{"given":"Androniki","family":"Menelaou","sequence":"first","affiliation":[{"name":"1 Department of Statistics, University of Oxford, Oxford OX1 3TG, United Kingdom and 2Wellcome Trust Centre for Human Genetics, Oxford, UK"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jonathan","family":"Marchini","sequence":"additional","affiliation":[{"name":"1 Department of Statistics, University of Oxford, Oxford OX1 3TG, United Kingdom and 2Wellcome Trust Centre for Human Genetics, Oxford, UK"},{"name":"1 Department of Statistics, University of Oxford, Oxford OX1 3TG, United Kingdom and 2Wellcome Trust Centre for Human Genetics, Oxford, UK"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"286","published-online":{"date-parts":[[2012,10,23]]},"reference":[{"key":"2023020303205293300_bts632-B1","doi-asserted-by":"crossref","DOI":"10.1002\/9780470015902.a0022496","article-title":"Haplotype sharing methods","volume-title":"Encyclopedia of Life Sciences","author":"Beckmann","year":"2010"},{"key":"2023020303205293300_bts632-B2","doi-asserted-by":"crossref","first-page":"210","DOI":"10.1016\/j.ajhg.2009.01.005","article-title":"A unified approach to genotype imputation and haplotype-phase inference for large data sets of trios and unrelated individuals","volume":"84","author":"Browning","year":"2009","journal-title":"Am. J. Hum. Genet."},{"key":"2023020303205293300_bts632-B3","doi-asserted-by":"crossref","first-page":"18","DOI":"10.1159\/000073729","article-title":"Detecting disease associations due to linkage disequilibrium using haplotype tags: a class of tests and the determinants of statistical power","volume":"56","author":"Chapman","year":"2003","journal-title":"Hum. Hered."},{"key":"2023020303205293300_bts632-B4","doi-asserted-by":"crossref","first-page":"1411","DOI":"10.1534\/genetics.110.114819","article-title":"Using environmental correlations to identify loci underlying local adaptation","volume":"185","author":"Coop","year":"2010","journal-title":"Genetics"},{"key":"2023020303205293300_bts632-B5","doi-asserted-by":"crossref","first-page":"179","DOI":"10.1038\/nmeth.1785","article-title":"A linear complexity phasing method for thousands of genomes","volume":"9","author":"Delaneau","year":"2011","journal-title":"Nat. Methods"},{"key":"2023020303205293300_bts632-B6","volume-title":"Matrix computations","author":"Golub","year":"1996","edition":"3rd"},{"key":"2023020303205293300_bts632-B7","doi-asserted-by":"crossref","first-page":"64","DOI":"10.1007\/978-3-642-29627-7_8","article-title":"Hap-seq: an optimal algorithm for haplotype phasing with imputation using sequencing data","volume":"7262","author":"He","year":"2012","journal-title":"Lect. Notes Comput. Sci."},{"key":"2023020303205293300_bts632-B8","doi-asserted-by":"crossref","first-page":"457","DOI":"10.1534\/g3.111.001198","article-title":"Genotype imputation with thousands of genomes","volume":"1","author":"Howie","year":"2011","journal-title":"G3 (Bethesda, Md.)"},{"key":"2023020303205293300_bts632-B9","doi-asserted-by":"crossref","first-page":"e1000529","DOI":"10.1371\/journal.pgen.1000529","article-title":"A flexible and accurate genotype imputation method for the next generation of genome-wide association studies","volume":"5","author":"Howie","year":"2009","journal-title":"PLoS. Genet."},{"key":"2023020303205293300_bts632-B10","doi-asserted-by":"crossref","first-page":"52","DOI":"10.1038\/nature09298","article-title":"Integrating common and rare genetic variation in diverse human populations","volume":"467","author":"International HapMap 3 Consortium (2010)","year":"2010","journal-title":"Nature"},{"key":"2023020303205293300_bts632-B11","doi-asserted-by":"crossref","first-page":"479","DOI":"10.1002\/gepi.20501","article-title":"Design of association studies with pooled or un-pooled next-generation sequencing data","volume":"34","author":"Kim","year":"2010","journal-title":"Genet. Epidemiol."},{"key":"2023020303205293300_bts632-B12","doi-asserted-by":"crossref","first-page":"1068","DOI":"10.1038\/ng.216","article-title":"Detection of sharing by descent, long-range phasing and haplotype imputation","volume":"40","author":"Kong","year":"2008","journal-title":"Nat. Genet."},{"key":"2023020303205293300_bts632-B13","doi-asserted-by":"crossref","first-page":"952","DOI":"10.1101\/gr.113084.110","article-title":"SNP detection and genotyping from low-coverage sequencing data on multiple diploid samples","volume":"21","author":"Le","year":"2011","journal-title":"Genome Res."},{"key":"2023020303205293300_bts632-B14","doi-asserted-by":"crossref","first-page":"2078","DOI":"10.1093\/bioinformatics\/btp352","article-title":"The sequence alignment\/map format and SAMtools","volume":"25","author":"Li","year":"2009","journal-title":"Bioinformatics"},{"key":"2023020303205293300_bts632-B15","doi-asserted-by":"crossref","first-page":"4384","DOI":"10.1093\/bioinformatics\/bti732","article-title":"Haplotype-based linkage disequilibrium mapping via direct data mining","volume":"21","author":"Li","year":"2005","journal-title":"Bioinformatics"},{"key":"2023020303205293300_bts632-B16","doi-asserted-by":"crossref","first-page":"940","DOI":"10.1101\/gr.117259.110","article-title":"Low-coverage sequencing: implications for design of complex trait association studies","volume":"21","author":"Li","year":"2011","journal-title":"Genome Res."},{"key":"2023020303205293300_bts632-B17","doi-asserted-by":"crossref","first-page":"816","DOI":"10.1002\/gepi.20533","article-title":"Mach: using sequence and genotype data to estimate haplotypes and unobserved genotypes","volume":"34","author":"Li","year":"2010","journal-title":"Genet. Epidemiol."},{"key":"2023020303205293300_bts632-B18","doi-asserted-by":"crossref","first-page":"747","DOI":"10.1038\/nature08494","article-title":"Finding the missing heritability of complex diseases","volume":"461","author":"Manolio","year":"2009","journal-title":"Nature"},{"key":"2023020303205293300_bts632-B19","doi-asserted-by":"crossref","first-page":"499","DOI":"10.1038\/nrg2796","article-title":"Genotype imputation for genome-wide association studies","volume":"11","author":"Marchini","year":"2010","journal-title":"Nat. Rev. Genet."},{"key":"2023020303205293300_bts632-B20","doi-asserted-by":"crossref","first-page":"906","DOI":"10.1038\/ng2088","article-title":"A new multipoint method for genome-wide association studies by imputation of genotypes","volume":"39","author":"Marchini","year":"2007","journal-title":"Nat. Genet."},{"key":"2023020303205293300_bts632-B21","doi-asserted-by":"crossref","first-page":"356","DOI":"10.1038\/nrg2344","article-title":"Genome-wide association studies for complex traits: consensus, uncertainty and challenges","volume":"9","author":"McCarthy","year":"2008","journal-title":"Nat. Rev. Genet."},{"key":"2023020303205293300_bts632-B22","doi-asserted-by":"crossref","first-page":"695","DOI":"10.1111\/1467-9868.00357","article-title":"Assessing population differentiation and isolation from single-nucleotide polymorphism data","volume":"64","author":"Nicholson","year":"2002","journal-title":"J. R. Stat. Soc. B"},{"key":"2023020303205293300_bts632-B23","doi-asserted-by":"crossref","first-page":"527","DOI":"10.1002\/gepi.21657","article-title":"Joint genotype calling with array and sequence data","volume":"36","author":"O\u2019Connell","year":"2012","journal-title":"Genet. Epidemiol."},{"key":"2023020303205293300_bts632-B24","doi-asserted-by":"crossref","first-page":"631","DOI":"10.1038\/ng.2283","article-title":"Extremely low-coverage sequencing and imputation increases power for genome-wide association studies","volume":"44","author":"Pasaniuc","year":"2012","journal-title":"Nat. Genet."},{"key":"2023020303205293300_bts632-B25","doi-asserted-by":"crossref","first-page":"1162","DOI":"10.1086\/379378","article-title":"A comparison of Bayesian methods for haplotype reconstruction from population genotype data","volume":"73","author":"Stephens","year":"2003","journal-title":"Am. J. Hum Genet."},{"key":"2023020303205293300_bts632-B26"},{"key":"2023020303205293300_bts632-B27","first-page":"195","article-title":"On the stability of inverse problems","volume":"39","author":"Tychonoff","year":"1943","journal-title":"Doklady. Akademii. Nauk. SSSR"},{"key":"2023020303205293300_bts632-B28","doi-asserted-by":"crossref","first-page":"1158","DOI":"10.1214\/10-AOAS338","article-title":"Using linear predictors to impute allele frequencies from summary or pooled genotype data","volume":"4","author":"Wen","year":"2010","journal-title":"Ann. Appl. Stat."}],"container-title":["Bioinformatics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/academic.oup.com\/bioinformatics\/article-pdf\/29\/1\/84\/49060693\/bioinformatics_29_1_84.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/academic.oup.com\/bioinformatics\/article-pdf\/29\/1\/84\/49060693\/bioinformatics_29_1_84.pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2023,2,3]],"date-time":"2023-02-03T03:22:43Z","timestamp":1675394563000},"score":1,"resource":{"primary":{"URL":"https:\/\/academic.oup.com\/bioinformatics\/article\/29\/1\/84\/272481"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2012,10,23]]},"references-count":28,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2013,1,1]]}},"URL":"https:\/\/doi.org\/10.1093\/bioinformatics\/bts632","relation":{},"ISSN":["1367-4811","1367-4803"],"issn-type":[{"value":"1367-4811","type":"electronic"},{"value":"1367-4803","type":"print"}],"subject":[],"published-other":{"date-parts":[[2013,1]]},"published":{"date-parts":[[2012,10,23]]}}}