{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,28]],"date-time":"2026-07-28T00:39:36Z","timestamp":1785199176654,"version":"3.55.0"},"reference-count":22,"publisher":"Oxford University Press (OUP)","issue":"21","content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2015,11,1]]},"abstract":"<jats:title>Abstract<\/jats:title>\n               <jats:p>Motivation: Next-generation sequencing produces vast amounts of data with errors that are difficult to distinguish from true biological variation when coverage is low.<\/jats:p>\n               <jats:p>Results: We demonstrate large reductions in error frequencies, especially for high-error-rate reads, by three independent means: (i) filtering reads according to their expected number of errors, (ii) assembling overlapping read pairs and (iii) for amplicon reads, by exploiting unique sequence abundances to perform error correction. We also show that most published paired read assemblers calculate incorrect posterior quality scores.<\/jats:p>\n               <jats:p>Availability and implementation: These methods are implemented in the USEARCH package. Binaries are freely available at http:\/\/drive5.com\/usearch.<\/jats:p>\n               <jats:p>Contact: \u00a0robert@drive5.com<\/jats:p>\n               <jats:p>Supplementary information: Supplementary data are available at Bioinformatics online.<\/jats:p>","DOI":"10.1093\/bioinformatics\/btv401","type":"journal-article","created":{"date-parts":[[2015,7,3]],"date-time":"2015-07-03T19:10:45Z","timestamp":1435950645000},"page":"3476-3482","source":"Crossref","is-referenced-by-count":1188,"title":["Error filtering, pair assembly and error correction for next-generation sequencing reads"],"prefix":"10.1093","volume":"31","author":[{"given":"Robert C.","family":"Edgar","sequence":"first","affiliation":[{"name":"1 Tiburon, CA 94920, USA and"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Henrik","family":"Flyvbjerg","sequence":"additional","affiliation":[{"name":"2 Department of Micro- and Nanotechnology, Technical University of Denmark, DK-2800 Lyngby, Denmark"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"286","published-online":{"date-parts":[[2015,7,2]]},"reference":[{"key":"2023020202323968200_btv401-B1","doi-asserted-by":"crossref","first-page":"57","DOI":"10.1038\/nmeth.2276","article-title":"Quality-filtering vastly improves diversity estimates from Illumina amplicon sequencing","volume":"10","author":"Bokulich","year":"2013","journal-title":"Nat. Methods"},{"key":"2023020202323968200_btv401-B2","doi-asserted-by":"crossref","first-page":"335","DOI":"10.1038\/nmeth.f.303","article-title":"QIIME allows analysis of high-throughput community sequencing data","volume":"7","author":"Caporaso","year":"2010","journal-title":"Nat. Methods"},{"key":"2023020202323968200_btv401-B3","doi-asserted-by":"crossref","first-page":"2460","DOI":"10.1093\/bioinformatics\/btq461","article-title":"Search and clustering orders of magnitude faster than BLAST","volume":"26","author":"Edgar","year":"2010","journal-title":"Bioinformatics"},{"key":"2023020202323968200_btv401-B4","doi-asserted-by":"crossref","first-page":"996","DOI":"10.1038\/nmeth.2604","article-title":"UPARSE: highly accurate OTU sequences from microbial amplicon reads","volume":"10","author":"Edgar","year":"2013","journal-title":"Nat. Methods"},{"key":"2023020202323968200_btv401-B5","doi-asserted-by":"crossref","first-page":"2194","DOI":"10.1093\/bioinformatics\/btr381","article-title":"UCHIME improves sensitivity and speed of chimera detection","volume":"27","author":"Edgar","year":"2011","journal-title":"Bioinformatics"},{"key":"2023020202323968200_btv401-B6","doi-asserted-by":"crossref","first-page":"759","DOI":"10.1111\/j.1755-0998.2011.03024.x","article-title":"Field guide to next-generation DNA sequencers","volume":"11","author":"Glenn","year":"2011","journal-title":"Mol. Ecol. Resour."},{"key":"2023020202323968200_btv401-B7","doi-asserted-by":"crossref","first-page":"494","DOI":"10.1101\/gr.112730.110","article-title":"Chimeric 16S rRNA sequence formation and detection in Sanger and 454-pyrosequenced PCR amplicons","author":"Haas","year":"2011","journal-title":"Genome Res."},{"key":"2023020202323968200_btv401-B8","doi-asserted-by":"crossref","first-page":"1889","DOI":"10.1111\/j.1462-2920.2010.02193.x","article-title":"Ironing out the wrinkles in the rare biosphere through improved OTU clustering","volume":"12","author":"Huse","year":"2010","journal-title":"Environ. Microbiol."},{"key":"2023020202323968200_btv401-B9","doi-asserted-by":"crossref","first-page":"171ra19","DOI":"10.1126\/scitranslmed.3004794","article-title":"Lineage structure of the human antibody repertoire in response to influenza vaccination","volume":"5","author":"Jiang","year":"2013","journal-title":"Sci. Transl. Med."},{"key":"2023020202323968200_btv401-B10","doi-asserted-by":"crossref","first-page":"5112","DOI":"10.1128\/AEM.01043-13","article-title":"Development of a dual-index sequencing strategy and curation pipeline for analyzing amplicon sequence data on the MiSeq illumina sequencing platform","volume":"79","author":"Kozich","year":"2013","journal-title":"Appl. Environ. Microbiol."},{"key":"2023020202323968200_btv401-B11","doi-asserted-by":"crossref","first-page":"1181","DOI":"10.2140\/pjm.1960.10.1181","article-title":"An approximation theorem for the Poisson binomial distribution","volume":"10","author":"Le Cam","year":"1960","journal-title":"Pacific J. Math."},{"key":"2023020202323968200_btv401-B12","doi-asserted-by":"crossref","first-page":"2870","DOI":"10.1093\/bioinformatics\/bts563","article-title":"COPE: An accurate k-mer-based pair-end reads connection tool to facilitate genome assembly","volume":"28","author":"Liu","year":"2012","journal-title":"Bioinformatics"},{"key":"2023020202323968200_btv401-B13","doi-asserted-by":"crossref","first-page":"2957","DOI":"10.1093\/bioinformatics\/btr507","article-title":"FLASH: fast length adjustment of short reads to improve genome assemblies","volume":"27","author":"Mago\u010d","year":"2011","journal-title":"Bioinformatics"},{"key":"2023020202323968200_btv401-B14","doi-asserted-by":"crossref","first-page":"31","DOI":"10.1186\/1471-2105-13-31","article-title":"PANDAseq: paired-end assembler for illumina sequences","volume":"13","author":"Masella","year":"2012","journal-title":"BMC Bioinformatics"},{"key":"2023020202323968200_btv401-B15","doi-asserted-by":"crossref","first-page":"639","DOI":"10.1038\/nmeth.1361","article-title":"Accurate determination of microbial diversity from 454 pyrosequencing data","volume":"6","author":"Quince","year":"2009","journal-title":"Nat. Methods"},{"key":"2023020202323968200_btv401-B16","doi-asserted-by":"crossref","first-page":"38","DOI":"10.1186\/1471-2105-12-38","article-title":"Removing noise from pyrosequenced amplicons","volume":"12","author":"Quince","year":"2011","journal-title":"BMC Bioinformatics"},{"key":"2023020202323968200_btv401-B17","doi-asserted-by":"crossref","first-page":"668","DOI":"10.1038\/nmeth0910-668b","article-title":"Rapidly denoising pyrosequencing amplicon reads by exploiting rank-abundance distributions","volume":"7","author":"Reeder","year":"2010","journal-title":"Nat. Methods"},{"key":"2023020202323968200_btv401-B18","doi-asserted-by":"crossref","first-page":"1198","DOI":"10.1038\/ismej.2013.227","article-title":"Host-specificity among abundant and rare taxa in the sponge microbiome","volume":"8","author":"Reveillaud","year":"2014","journal-title":"ISME J."},{"key":"2023020202323968200_btv401-B19","doi-asserted-by":"crossref","DOI":"10.1371\/journal.pone.0011840","article-title":"Unlocking short read sequencing for metagenomics","volume":"5","author":"Rodrigue","year":"2010","journal-title":"PLoS One"},{"key":"2023020202323968200_btv401-B20","doi-asserted-by":"crossref","first-page":"e27310","DOI":"10.1371\/journal.pone.0027310","article-title":"Reducing the effects of PCR amplification and sequencing artifacts on 16S rRNA-based studies","volume":"6","author":"Schloss","year":"2011","journal-title":"PLoS One"},{"key":"2023020202323968200_btv401-B21","first-page":"295","article-title":"On the number of successes in independent trials","volume":"3","author":"Wang","year":"1993","journal-title":"Stat. Sin."},{"key":"2023020202323968200_btv401-B22","doi-asserted-by":"crossref","first-page":"614","DOI":"10.1093\/bioinformatics\/btt593","article-title":"PEAR: a fast and accurate Illumina paired-end read merger","volume":"30","author":"Zhang","year":"2014","journal-title":"Bioinformatics"}],"container-title":["Bioinformatics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/academic.oup.com\/bioinformatics\/article-pdf\/31\/21\/3476\/49035499\/bioinformatics_31_21_3476.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/academic.oup.com\/bioinformatics\/article-pdf\/31\/21\/3476\/49035499\/bioinformatics_31_21_3476.pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2023,2,2]],"date-time":"2023-02-02T03:51:40Z","timestamp":1675309900000},"score":1,"resource":{"primary":{"URL":"https:\/\/academic.oup.com\/bioinformatics\/article\/31\/21\/3476\/194979"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2015,7,2]]},"references-count":22,"journal-issue":{"issue":"21","published-print":{"date-parts":[[2015,11,1]]}},"URL":"https:\/\/doi.org\/10.1093\/bioinformatics\/btv401","relation":{},"ISSN":["1367-4811","1367-4803"],"issn-type":[{"value":"1367-4811","type":"electronic"},{"value":"1367-4803","type":"print"}],"subject":[],"published-other":{"date-parts":[[2015,11,1]]},"published":{"date-parts":[[2015,7,2]]}}}