{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,14]],"date-time":"2026-05-14T14:53:56Z","timestamp":1778770436569,"version":"3.51.4"},"reference-count":25,"publisher":"Oxford University Press (OUP)","issue":"6","content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2014,3,15]]},"abstract":"<jats:title>Abstract<\/jats:title><jats:p>Motivation: ChIP-seq technology enables investigators to study genome-wide binding of transcription factors and mapping of epigenomic marks. Although the availability of basic analysis tools for ChIP-seq data is rapidly increasing, there has not been much progress on the related design issues. A challenging question for designing a ChIP-seq experiment is how deeply should the ChIP and the control samples be sequenced? The answer depends on multiple factors some of which can be set by the experimenter based on pilot\/preliminary data. The sequencing depth of a ChIP-seq experiment is one of the key factors that determine whether all the underlying targets (e.g. binding locations or epigenomic profiles) can be identified with a targeted power.<\/jats:p><jats:p>Results: We developed a statistical framework named CSSP (ChIP-seq Statistical Power) for power calculations in ChIP-seq experiments by considering a local Poisson model, which is commonly adopted by many peak callers. Evaluations with simulations and data-driven computational experiments demonstrate that this framework can reliably estimate the power of a ChIP-seq experiment at different sequencing depths based on pilot data. Furthermore, it provides an analytical approach for calculating the required depth for a targeted power while controlling the false discovery rate at a user-specified level. Hence, our results enable researchers to use their own or publicly available data for determining required sequencing depths of their ChIP-seq experiments and potentially make better use of the multiplexing functionality of the sequencers. Evaluation of power for multiple public ChIP-seq datasets indicate that, currently, typical ChIP-seq studies are powered well for detecting large fold changes of ChIP enrichment over the control sample, but they have considerably less power for detecting smaller fold changes.<\/jats:p><jats:p>Availability: Available at www.stat.wisc.edu\/\u223czuo\/CSSP.<\/jats:p><jats:p>Contact: \u00a0keles@stat.wisc.edu<\/jats:p><jats:p>Supplementary information: \u00a0Supplementary data are available at Bioinformatics online.<\/jats:p>","DOI":"10.1093\/bioinformatics\/btt200","type":"journal-article","created":{"date-parts":[[2013,5,11]],"date-time":"2013-05-11T01:47:48Z","timestamp":1368236868000},"page":"753-760","source":"Crossref","is-referenced-by-count":15,"title":["A statistical framework for power calculations in ChIP-seq experiments"],"prefix":"10.1093","volume":"30","author":[{"given":"Chandler","family":"Zuo","sequence":"first","affiliation":[{"name":"1 Department of Statistics, and 2Department of Biostatistics and Medical Informatics, 1300 University Avenue, Madison, WI 53706, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"S\u00fcnd\u00fcz","family":"Kele\u015f","sequence":"additional","affiliation":[{"name":"1 Department of Statistics, and 2Department of Biostatistics and Medical Informatics, 1300 University Avenue, Madison, WI 53706, USA"},{"name":"1 Department of Statistics, and 2Department of Biostatistics and Medical Informatics, 1300 University Avenue, Madison, WI 53706, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"286","published-online":{"date-parts":[[2013,5,10]]},"reference":[{"key":"2023012710442881400_btt200-B1","doi-asserted-by":"crossref","first-page":"289","DOI":"10.1111\/j.2517-6161.1995.tb02031.x","article-title":"Controlling the false discovery rate: a practical and powerful approach to multiple testing","volume":"57","author":"Benjamini","year":"1995","journal-title":"J. R. Stat. Soc. Series B Methodol."},{"key":"2023012710442881400_btt200-B2","doi-asserted-by":"crossref","first-page":"609","DOI":"10.1038\/nmeth.1985","article-title":"Systematic evaluation of factors influencing ChIP-seq fidelity","volume":"9","author":"Chen","year":"2012","journal-title":"Nat. Methods"},{"key":"2023012710442881400_btt200-B3","doi-asserted-by":"crossref","first-page":"1775","DOI":"10.1126\/science.1196914","article-title":"Integrative analysis of the caenorhabditis elegans genome by the modENCODE project","volume":"330","author":"Gerstein","year":"2010","journal-title":"Science"},{"key":"2023012710442881400_btt200-B4","doi-asserted-by":"crossref","first-page":"1760","DOI":"10.1101\/gr.135350.111","article-title":"GENCODE: the reference human genome annotation for the ENCODE project","volume":"22","author":"Harrow","year":"2012","journal-title":"Genome Res."},{"key":"2023012710442881400_btt200-B5","doi-asserted-by":"crossref","first-page":"134","DOI":"10.1186\/1471-2164-12-134","article-title":"ChIP-chip versus ChIP-seq: Lessons for experimental design and data analysis","volume":"12","author":"Ho","year":"2011","journal-title":"BMC Genomics"},{"key":"2023012710442881400_btt200-B6","doi-asserted-by":"crossref","first-page":"1293","DOI":"10.1038\/nbt.1505","article-title":"An integrated software system for analyzing ChIP-chip and ChIP-seq data","volume":"26","author":"Ji","year":"2008","journal-title":"Nat. Biot."},{"key":"2023012710442881400_btt200-B7","doi-asserted-by":"crossref","first-page":"232","DOI":"10.1126\/science.1183621","article-title":"Variation in transcription factor binding among humans","volume":"328","author":"Kasowski","year":"2010","journal-title":"Science"},{"key":"2023012710442881400_btt200-B8","doi-asserted-by":"crossref","first-page":"1351","DOI":"10.1038\/nbt.1508","article-title":"Design and analysis of ChIP-seq experiments for DNA-binding proteins","volume":"6","author":"Kharchenko","year":"2008","journal-title":"Nat. Biotechnol."},{"key":"2023012710442881400_btt200-B9","doi-asserted-by":"crossref","first-page":"891","DOI":"10.1198\/jasa.2011.ap09706","article-title":"A statistical framework for the analysis of ChIP-Seq data","volume":"106","author":"Kuan","year":"2011","journal-title":"J. Am. Stat. Assoc."},{"key":"2023012710442881400_btt200-B10","doi-asserted-by":"crossref","first-page":"199","DOI":"10.1186\/1471-2105-13-199","article-title":"Normalization of ChIP-seq data with control","volume":"13","author":"Liang","year":"2012","journal-title":"BMC Bioinformatics"},{"key":"2023012710442881400_btt200-B11","doi-asserted-by":"crossref","first-page":"235","DOI":"10.1126\/science.1184655","article-title":"Heritable individual-specific and allele-specific chromatin signatures in humans","volume":"328","author":"McDaniell","year":"2010","journal-title":"Science"},{"key":"2023012710442881400_btt200-B25","doi-asserted-by":"crossref","first-page":"e1003565","DOI":"10.1371\/journal.pgen.1003565","article-title":"Genome-scale analysis of Escherichia coli FNR reveals complex features of transcription factor binding","volume":"9","author":"Myers","year":"2013","journal-title":"PLoS Genet."},{"key":"2023012710442881400_btt200-B12","first-page":"21","article-title":"A Users Guide to the Encyclopedia of DNA Elements (ENCODE)","volume":"9","author":"Myers","year":"2011","journal-title":"PLoS Biol."},{"key":"2023012710442881400_btt200-B13","doi-asserted-by":"crossref","first-page":"523","DOI":"10.1186\/1471-2105-9-523","article-title":"Empirical methods for controlling false positives and estimating confidence in ChIP-Seq peaks","volume":"9","author":"Nix","year":"2008","journal-title":"BMC Bioinformatics"},{"key":"2023012710442881400_btt200-B14","doi-asserted-by":"crossref","first-page":"616","DOI":"10.1080\/01621459.1980.10477522","article-title":"Minimum distance and robust estimation","volume":"75","author":"Parr","year":"1980","journal-title":"J. Am. Stat. Assoc."},{"key":"2023012710442881400_btt200-B15","doi-asserted-by":"crossref","first-page":"R67","DOI":"10.1186\/gb-2011-12-7-r67","article-title":"ZINBA integrates local covariates with DNA-seq data to identify broad and narrow regions of enrichment, even within amplified genomic regions","volume":"12","author":"Rashid","year":"2011","journal-title":"Genome Biol."},{"key":"2023012710442881400_btt200-B16","doi-asserted-by":"crossref","first-page":"1787","DOI":"10.1126\/science.1198374","article-title":"Identification of functional elements and regulatory circuits by Drosophila modENCODE","volume":"330","author":"Roy","year":"2010","journal-title":"Science"},{"key":"2023012710442881400_btt200-B17","doi-asserted-by":"crossref","first-page":"66","DOI":"10.1038\/nbt.1518","article-title":"PeakSeq enables systematic scoring of ChIP-Seq experiments relative to controls","volume":"27","author":"Rozowsky","year":"2009","journal-title":"Nat. Biotechnol."},{"key":"2023012710442881400_btt200-B18","doi-asserted-by":"crossref","first-page":"461","DOI":"10.1214\/aos\/1176344136","article-title":"Estimating the dimension of a model","volume":"6","author":"Schwarz","year":"1978","journal-title":"Ann. Stat."},{"key":"2023012710442881400_btt200-B19","doi-asserted-by":"crossref","first-page":"57","DOI":"10.1038\/nature11247","article-title":"An integrated encyclopedia of DNA elements in the human genome","volume":"489","author":"The ENCODE Project Consortium","year":"2012","journal-title":"Nature"},{"key":"2023012710442881400_btt200-B20","doi-asserted-by":"crossref","first-page":"e11471","DOI":"10.1371\/journal.pone.0011471","article-title":"Evaluation of algorithm performance in ChIP-seq peak detection","volume":"5","author":"Wilbanks","year":"2010","journal-title":"PLoS One"},{"key":"2023012710442881400_btt200-B21","doi-asserted-by":"crossref","first-page":"1659","DOI":"10.1101\/gr.125088.111","article-title":"Dynamics of the epigenetic landscape during erythroid differentiation after GATA1 restoration","volume":"21","author":"Wu","year":"2011","journal-title":"Genome Res."},{"key":"2023012710442881400_btt200-B22","doi-asserted-by":"crossref","first-page":"1199","DOI":"10.1093\/bioinformatics\/btq128","article-title":"A signal-noise model for significance analysis of ChIP-seq with negative control","volume":"26","author":"Xu","year":"2010","journal-title":"Bioinformatics"},{"key":"2023012710442881400_btt200-B23","doi-asserted-by":"crossref","first-page":"151163","DOI":"10.1111\/j.1541-0420.2010.01441.x","article-title":"Probabilistic inference for ChIP-seq","volume":"67","author":"Zhang","year":"2011","journal-title":"Biometrics"},{"key":"2023012710442881400_btt200-B24","doi-asserted-by":"crossref","first-page":"R137","DOI":"10.1186\/gb-2008-9-9-r137","article-title":"Model-based analysis of ChIP-Seq (MACS)","volume":"9","author":"Zhang","year":"2008","journal-title":"Genome Biol."}],"container-title":["Bioinformatics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/academic.oup.com\/bioinformatics\/article-pdf\/30\/6\/753\/48921816\/bioinformatics_30_6_753.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/academic.oup.com\/bioinformatics\/article-pdf\/30\/6\/753\/48921816\/bioinformatics_30_6_753.pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,5,10]],"date-time":"2024-05-10T13:46:35Z","timestamp":1715348795000},"score":1,"resource":{"primary":{"URL":"https:\/\/academic.oup.com\/bioinformatics\/article\/30\/6\/753\/285195"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2013,5,10]]},"references-count":25,"journal-issue":{"issue":"6","published-print":{"date-parts":[[2014,3,15]]}},"URL":"https:\/\/doi.org\/10.1093\/bioinformatics\/btt200","relation":{},"ISSN":["1367-4811","1367-4803"],"issn-type":[{"value":"1367-4811","type":"electronic"},{"value":"1367-4803","type":"print"}],"subject":[],"published-other":{"date-parts":[[2014,3,15]]},"published":{"date-parts":[[2013,5,10]]}}}