{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2023,9,8]],"date-time":"2023-09-08T13:24:59Z","timestamp":1694179499233},"reference-count":65,"publisher":"Oxford University Press (OUP)","issue":"5","content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2008,3,1]]},"abstract":"<jats:title>Abstract<\/jats:title>\n               <jats:p>Motivation: Accurate automatic assignment of protein functions remains a challenge for genome annotation. We have developed and compared the automatic annotation of four bacterial genomes employing a 5-fold cross-validation procedure and several machine learning methods.<\/jats:p>\n               <jats:p>Results: The analyzed genomes were manually annotated with FunCat categories in MIPS providing a gold standard. Features describing a pair of sequences rather than each sequence alone were used. The descriptors were derived from sequence alignment scores, InterPro domains, synteny information, sequence length and calculated protein properties. Following training we scored all pairs from the validation sets, selected a pair with the highest predicted score and annotated the target protein with functional categories of the prototype protein. The data integration using machine-learning methods provided significantly higher annotation accuracy compared to the use of individual descriptors alone. The neural network approach showed the best performance. The descriptors derived from the InterPro domains and sequence similarity provided the highest contribution to the method performance. The predicted annotation scores allow differentiation of reliable versus non-reliable annotations. The developed approach was applied to annotate the protein sequences from 180 complete bacterial genomes.<\/jats:p>\n               <jats:p>Availability: The FUNcat Annotation Tool (FUNAT) is available on-line as Web Services at http:\/\/mips.gsf.de\/proj\/funat<\/jats:p>\n               <jats:p>Contact: \u00a0i.tetko@gsf.de<\/jats:p>\n               <jats:p>Supplementary information: Supplementary data are available at Bioinformatics online.<\/jats:p>","DOI":"10.1093\/bioinformatics\/btm633","type":"journal-article","created":{"date-parts":[[2008,1,4]],"date-time":"2008-01-04T01:13:32Z","timestamp":1199409212000},"page":"621-628","source":"Crossref","is-referenced-by-count":7,"title":["Beyond the \u2018best\u2019 match: machine learning annotation of protein sequences by integration of different sources of information"],"prefix":"10.1093","volume":"24","author":[{"given":"Igor V.","family":"Tetko","sequence":"first","affiliation":[{"name":"1 Helmholtz Zentrum M\u00fcnchen - German Research Center for Environmental Health (GmbH), Institute of Bioinformatics and Systems Biology, Ingolst\u00e4dter Landstra\u00dfe 1, 85764, Neuherberg and 2Department of Genome-Oriented Bioinformatics, Wissenschaftszentrum Weihenstephan, Technische Universit\u00e4t M\u00fcnchen 85350 Freising, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Igor V.","family":"Rodchenkov","sequence":"additional","affiliation":[{"name":"1 Helmholtz Zentrum M\u00fcnchen - German Research Center for Environmental Health (GmbH), Institute of Bioinformatics and Systems Biology, Ingolst\u00e4dter Landstra\u00dfe 1, 85764, Neuherberg and 2Department of Genome-Oriented Bioinformatics, Wissenschaftszentrum Weihenstephan, Technische Universit\u00e4t M\u00fcnchen 85350 Freising, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Mathias C.","family":"Walter","sequence":"additional","affiliation":[{"name":"1 Helmholtz Zentrum M\u00fcnchen - German Research Center for Environmental Health (GmbH), Institute of Bioinformatics and Systems Biology, Ingolst\u00e4dter Landstra\u00dfe 1, 85764, Neuherberg and 2Department of Genome-Oriented Bioinformatics, Wissenschaftszentrum Weihenstephan, Technische Universit\u00e4t M\u00fcnchen 85350 Freising, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Thomas","family":"Rattei","sequence":"additional","affiliation":[{"name":"1 Helmholtz Zentrum M\u00fcnchen - German Research Center for Environmental Health (GmbH), Institute of Bioinformatics and Systems Biology, Ingolst\u00e4dter Landstra\u00dfe 1, 85764, Neuherberg and 2Department of Genome-Oriented Bioinformatics, Wissenschaftszentrum Weihenstephan, Technische Universit\u00e4t M\u00fcnchen 85350 Freising, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Hans-Werner","family":"Mewes","sequence":"additional","affiliation":[{"name":"1 Helmholtz Zentrum M\u00fcnchen - German Research Center for Environmental Health (GmbH), Institute of Bioinformatics and Systems Biology, Ingolst\u00e4dter Landstra\u00dfe 1, 85764, Neuherberg and 2Department of Genome-Oriented Bioinformatics, Wissenschaftszentrum Weihenstephan, Technische Universit\u00e4t M\u00fcnchen 85350 Freising, Germany"},{"name":"1 Helmholtz Zentrum M\u00fcnchen - German Research Center for Environmental Health (GmbH), Institute of Bioinformatics and Systems Biology, Ingolst\u00e4dter Landstra\u00dfe 1, 85764, Neuherberg and 2Department of Genome-Oriented Bioinformatics, Wissenschaftszentrum Weihenstephan, Technische Universit\u00e4t M\u00fcnchen 85350 Freising, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"286","published-online":{"date-parts":[[2008,1,3]]},"reference":[{"key":"2023020210104500100_B1","doi-asserted-by":"crossref","first-page":"683","DOI":"10.1002\/prot.10449","article-title":"Automatic annotation of protein function based on family identification","volume":"53","author":"Abascal","year":"2003","journal-title":"Proteins"},{"key":"2023020210104500100_B2","doi-asserted-by":"crossref","first-page":"3389","DOI":"10.1093\/nar\/25.17.3389","article-title":"Gapped BLAST and PSI-BLAST: a new generation of protein database search programs","volume":"25","author":"Altschul","year":"1997","journal-title":"Nucleic Acids Res"},{"key":"2023020210104500100_B3","doi-asserted-by":"crossref","first-page":"391","DOI":"10.1093\/bioinformatics\/15.5.391","article-title":"Automated genome sequence analysis and annotation","volume":"15","author":"Andrade","year":"1999","journal-title":"Bioinformatics"},{"key":"2023020210104500100_B4","doi-asserted-by":"crossref","first-page":"ii42","DOI":"10.1093\/bioinformatics\/bti1107","article-title":"SIMAP\u2014The similarity matrix of proteins","volume":"21","author":"Arnold","year":"2005","journal-title":"Bioinformatics"},{"key":"2023020210104500100_B5","doi-asserted-by":"crossref","first-page":"25","DOI":"10.1038\/75556","article-title":"Gene ontology: tool for the unification of biology. The Gene Ontology Consortium","volume":"25","author":"Ashburner","year":"2000","journal-title":"Nat. Genet"},{"key":"2023020210104500100_B6","doi-asserted-by":"crossref","DOI":"10.1109\/ICDMW.2006.130","article-title":"Predictive integration of Gene Ontology-driven similarity and functional interactions","author":"Azuaje","year":"2006"},{"key":"2023020210104500100_B7","doi-asserted-by":"crossref","first-page":"D154","DOI":"10.1093\/nar\/gki070","article-title":"The universal protein resource (UniProt)","volume":"33","author":"Bairoch","year":"2005","journal-title":"Nucleic Acids Res"},{"key":"2023020210104500100_B8","doi-asserted-by":"crossref","first-page":"830","DOI":"10.1093\/bioinformatics\/btk048","article-title":"Hierarchical multi-label prediction of gene function","volume":"22","author":"Barutcuoglu","year":"2006","journal-title":"Bioinformatics"},{"key":"2023020210104500100_B9","doi-asserted-by":"crossref","first-page":"783","DOI":"10.1016\/j.jmb.2004.05.028","article-title":"Improved prediction of signal peptides: SignalP 3.0","volume":"340","author":"Bendtsen","year":"2004","journal-title":"J. Mol. Biol"},{"key":"2023020210104500100_B10","doi-asserted-by":"crossref","first-page":"285","DOI":"10.1093\/bib\/3.3.285","article-title":"Applications of interPro in protein annotation and genome analysis","volume":"3","author":"Biswas","year":"2002","journal-title":"Brief Bioinform"},{"key":"2023020210104500100_B11","article-title":"LIBSVM: a library for support vector machines","author":"Chang","year":"2001"},{"key":"2023020210104500100_B12","doi-asserted-by":"crossref","first-page":"1523","DOI":"10.1109\/TIT.2005.844059","article-title":"Clustering by compression","volume":"51","author":"Cilibrasi","year":"2005","journal-title":"IEEE Trans. Inf. Theory"},{"key":"2023020210104500100_B13","doi-asserted-by":"crossref","first-page":"1130","DOI":"10.1093\/bioinformatics\/btl051","article-title":"Functional bioinformatics for Arabidopsis thaliana","volume":"22","author":"Clare","year":"2006","journal-title":"Bioinformatics"},{"key":"2023020210104500100_B14","doi-asserted-by":"crossref","first-page":"II42","DOI":"10.1093\/bioinformatics\/btg1058","article-title":"Predicting gene function in Saccharomyces cerevisiae","volume":"19","author":"Clare","year":"2003","journal-title":"Bioinformatics"},{"key":"2023020210104500100_B15","doi-asserted-by":"crossref","first-page":"1575","DOI":"10.1093\/nar\/30.7.1575","article-title":"An efficient algorithm for large-scale detection of protein families","volume":"30","author":"Enright","year":"2002","journal-title":"Nucleic Acids Res"},{"key":"2023020210104500100_B16","doi-asserted-by":"crossref","first-page":"225","DOI":"10.1093\/bib\/bbl004","article-title":"Automated protein function prediction\u2013the genomic challenge","volume":"7","author":"Friedberg","year":"2006","journal-title":"Brief Bioinform"},{"key":"2023020210104500100_B17","doi-asserted-by":"crossref","first-page":"3448","DOI":"10.1021\/cr068303k","article-title":"Protein annotation at genomic scale: the current status","volume":"107","author":"Frishman","year":"2007","journal-title":"Chem. Rev"},{"key":"2023020210104500100_B18","doi-asserted-by":"crossref","first-page":"329","DOI":"10.1002\/(SICI)1097-0134(199703)27:3<329::AID-PROT1>3.0.CO;2-8","article-title":"Seventy-five percent accuracy in protein secondary structure prediction","volume":"27","author":"Frishman","year":"1997","journal-title":"Proteins"},{"key":"2023020210104500100_B19","doi-asserted-by":"crossref","first-page":"1257","DOI":"10.1016\/S0022-2836(02)00379-0","article-title":"Prediction of human protein function from post-translational modifications and localization features","volume":"319","author":"Jensen","year":"2002","journal-title":"J. Mol. Biol"},{"key":"2023020210104500100_B20","doi-asserted-by":"crossref","first-page":"635","DOI":"10.1093\/bioinformatics\/btg036","article-title":"Prediction of human protein function according to Gene Ontology categories","volume":"19","author":"Jensen","year":"2003","journal-title":"Bioinformatics"},{"key":"2023020210104500100_B21","doi-asserted-by":"crossref","first-page":"196","DOI":"10.1186\/1471-2105-5-196","article-title":"A functional hierarchical organization of the protein sequence space","volume":"5","author":"Kaplan","year":"2004","journal-title":"BMC Bioinformatics"},{"key":"2023020210104500100_B22","doi-asserted-by":"crossref","first-page":"407","DOI":"10.1093\/bioinformatics\/bti806","article-title":"Application of compression-based distance measures to protein sequence classification: a methodological study","volume":"22","author":"Kocsor","year":"2006","journal-title":"Bioinformatics"},{"key":"2023020210104500100_B23","doi-asserted-by":"crossref","first-page":"639","DOI":"10.1006\/jmbi.2001.4701","article-title":"SNAPping up functionally related genes based on context information: a colinearity-free approach","volume":"311","author":"Kolesov","year":"2001","journal-title":"J. Mol. Biol"},{"key":"2023020210104500100_B24","doi-asserted-by":"crossref","first-page":"1066","DOI":"10.1093\/bioinformatics\/bth039","article-title":"Statistically rigorous automated protein annotation","volume":"20","author":"Krebs","year":"2004","journal-title":"Bioinformatics"},{"key":"2023020210104500100_B25","doi-asserted-by":"crossref","first-page":"920","DOI":"10.1093\/bioinformatics\/17.10.920","article-title":"Automatic rule generation for protein annotation with the C4.5 data mining algorithm applied on SWISS-PROT","volume":"17","author":"Kretschmann","year":"2001","journal-title":"Bioinformatics"},{"key":"2023020210104500100_B26","doi-asserted-by":"crossref","first-page":"567","DOI":"10.1006\/jmbi.2000.4315","article-title":"Predicting transmembrane protein topology with a hidden Markov model: application to complete genomes","volume":"305","author":"Krogh","year":"2001","journal-title":"J. Mol. Biol"},{"key":"2023020210104500100_B27","doi-asserted-by":"crossref","first-page":"2626","DOI":"10.1093\/bioinformatics\/bth294","article-title":"A statistical framework for genomic data fusion","volume":"20","author":"Lanckriet","year":"2004","journal-title":"Bioinformatics"},{"key":"2023020210104500100_B28","first-page":"598","article-title":"Optimal Brain Damage","volume-title":"Advances in Neural Processing Systems II (NIPS*2).","author":"LeCun","year":"1990"},{"key":"2023020210104500100_B29","doi-asserted-by":"crossref","first-page":"302","DOI":"10.1186\/1471-2105-6-302","article-title":"Probabilistic annotation of protein sequences based on functional classifications","volume":"6","author":"Levy","year":"2005","journal-title":"BMC Bioinformatics"},{"key":"2023020210104500100_B30","first-page":"296","article-title":"An information-theoretic definition of similarity","author":"Lin","year":"1998"},{"key":"2023020210104500100_B31","doi-asserted-by":"crossref","first-page":"3701","DOI":"10.1093\/nar\/gkg519","article-title":"GlobPlot: Exploring protein sequences for globularity and disorder","volume":"31","author":"Linding","year":"2003","journal-title":"Nucleic Acids Res"},{"key":"2023020210104500100_B32","doi-asserted-by":"crossref","first-page":"513","DOI":"10.1016\/S0076-6879(96)66032-7","article-title":"Prediction and analysis of coiled-coil structures","volume":"266","author":"Lupas","year":"1996","journal-title":"Methods Enzymol"},{"key":"2023020210104500100_B33","doi-asserted-by":"crossref","first-page":"83","DOI":"10.1038\/47048","article-title":"A combined algorithm for genome-wide prediction of protein function","volume":"402","author":"Marcotte","year":"1999","journal-title":"Nature"},{"key":"2023020210104500100_B34","doi-asserted-by":"crossref","first-page":"1703","DOI":"10.1101\/gr.192502","article-title":"Systematic learning of gene functional classes from DNA array expression data by using multilayer perceptrons","volume":"12","author":"Mateos","year":"2002","journal-title":"Genome Res"},{"key":"2023020210104500100_B35","doi-asserted-by":"crossref","first-page":"D226","DOI":"10.1093\/nar\/gki030","article-title":"The SYSTERS protein family database in 2005","volume":"33","author":"Meinel","year":"2005","journal-title":"Nucleic Acids Res"},{"key":"2023020210104500100_B36","doi-asserted-by":"crossref","first-page":"7","DOI":"10.1038\/387s007","article-title":"Overview of the yeast genome","volume":"387","author":"Mewes","year":"1997","journal-title":"Nature"},{"key":"2023020210104500100_B37","doi-asserted-by":"crossref","first-page":"44","DOI":"10.1093\/nar\/27.1.44","article-title":"MIPS: a database for genomes and protein sequences","volume":"27","author":"Mewes","year":"1999","journal-title":"Nucleic Acids Res"},{"key":"2023020210104500100_B38","doi-asserted-by":"crossref","first-page":"D224","DOI":"10.1093\/nar\/gkl841","article-title":"New developments in the interPro database","volume":"35","author":"Mulder","year":"2007","journal-title":"Nucleic Acids Res"},{"key":"2023020210104500100_B39","doi-asserted-by":"crossref","first-page":"34","DOI":"10.1016\/S0968-0004(98)01336-X","article-title":"PSORT: a program for detecting sorting signals in proteins and predicting their subcellular localization","volume":"24","author":"Nakai","year":"1999","journal-title":"Trends Biochem. Sci"},{"key":"2023020210104500100_B40","doi-asserted-by":"crossref","first-page":"3","DOI":"10.1093\/protein\/12.1.3","article-title":"Machine learning approaches for the prediction of signal peptides and other protein sorting signals","volume":"12","author":"Nielsen","year":"1999","journal-title":"Protein Eng"},{"key":"2023020210104500100_B41","doi-asserted-by":"crossref","first-page":"1571","DOI":"10.1093\/nar\/gkj515","article-title":"Spectral clustering of protein sequences","volume":"34","author":"Paccanaro","year":"2006","journal-title":"Nucleic Acids Res"},{"key":"2023020210104500100_B42","doi-asserted-by":"crossref","first-page":"227","DOI":"10.1016\/S0076-6879(96)66017-0","article-title":"Effective protein sequence comparison","volume":"266","author":"Pearson","year":"1996","journal-title":"Methods Enzymol"},{"key":"2023020210104500100_B43","doi-asserted-by":"crossref","first-page":"D289","DOI":"10.1093\/nar\/gkm963","article-title":"SIMAP structuring the network of protein similarities","volume":"36","author":"Rattei","year":"2008","journal-title":"Nucleic Acids Res"},{"key":"2023020210104500100_B44","doi-asserted-by":"crossref","first-page":"1041","DOI":"10.1006\/jmbi.2000.5197","article-title":"Automatic clustering of orthologs and in-paralogs from pairwise species comparisons","volume":"314","author":"Remm","year":"2001","journal-title":"J. Mol. Biol"},{"key":"2023020210104500100_B45","first-page":"448","article-title":"Using information content to evaluate semantic similarity in a taxonomy","author":"Resnik","year":"1995"},{"key":"2023020210104500100_B46","doi-asserted-by":"crossref","first-page":"D354","DOI":"10.1093\/nar\/gkl1005","article-title":"PEDANT genome database: 10 years online","volume":"35","author":"Riley","year":"2007","journal-title":"Nucleic Acids Res"},{"key":"2023020210104500100_B47","doi-asserted-by":"crossref","first-page":"145","DOI":"10.1016\/j.ddtec.2006.06.011","article-title":"Prediction and Classification of Protein Functions","volume":"3","author":"Ruepp","year":"2006","journal-title":"Drug Discov. Today: Tech"},{"key":"2023020210104500100_B48","doi-asserted-by":"crossref","first-page":"5539","DOI":"10.1093\/nar\/gkh894","article-title":"The FunCat, a functional annotation scheme for systematic classification of proteins from whole genomes","volume":"32","author":"Ruepp","year":"2004","journal-title":"Nucleic Acids Res"},{"key":"2023020210104500100_B49","doi-asserted-by":"crossref","first-page":"195","DOI":"10.1016\/0022-2836(81)90087-5","article-title":"Identification of common molecular subsequences","volume":"147","author":"Smith","year":"1981","journal-title":"J. Mol. Biol"},{"key":"2023020210104500100_B50","doi-asserted-by":"crossref","first-page":"405","DOI":"10.1002\/(SICI)1097-0134(199707)28:3<405::AID-PROT10>3.0.CO;2-L","article-title":"Pfam: a comprehensive database of protein domain families based on seed alignments","volume":"28","author":"Sonnhammer","year":"1997","journal-title":"Proteins"},{"key":"2023020210104500100_B51","doi-asserted-by":"crossref","first-page":"187","DOI":"10.1023\/A:1019903710291","article-title":"Associative neural network","volume":"16","author":"Tetko","year":"2002","journal-title":"Neural Process. Lett"},{"key":"2023020210104500100_B52","doi-asserted-by":"crossref","first-page":"717","DOI":"10.1021\/ci010379o","article-title":"Neural network studies. 4. Introduction to associative neural networks","volume":"42","author":"Tetko","year":"2002","journal-title":"J. Chem. Inf. Comput. Sci"},{"key":"2023020210104500100_B53","doi-asserted-by":"crossref","first-page":"2520","DOI":"10.1093\/bioinformatics\/bti380","article-title":"MIPS bacterial genomes functional annotation benchmark dataset","volume":"21","author":"Tetko","year":"2005","journal-title":"Bioinformatics"},{"key":"2023020210104500100_B54","doi-asserted-by":"crossref","first-page":"82","DOI":"10.1186\/1471-2105-6-82","article-title":"Super paramagnetic clustering of protein sequences","volume":"6","author":"Tetko","year":"2005","journal-title":"BMC Bioinformatics"},{"key":"2023020210104500100_B55","doi-asserted-by":"crossref","first-page":"453","DOI":"10.1007\/s10822-005-8694-y","article-title":"Virtual computational chemistry laboratory - design and description","volume":"19","author":"Tetko","year":"2005","journal-title":"J. Comput.-Aided Mol. Des"},{"key":"2023020210104500100_B56","doi-asserted-by":"crossref","first-page":"808","DOI":"10.1021\/ci0504216","article-title":"Benchmarking of linear and nonlinear approaches for quantitative structure-property relationship studies of metal complexation with ionophores","volume":"46","author":"Tetko","year":"2006","journal-title":"J. Chem. Inf. Model"},{"key":"2023020210104500100_B57","doi-asserted-by":"crossref","first-page":"794","DOI":"10.1021\/ci950204c","article-title":"Neural network studies. 2. Variable selection","volume":"36","author":"Tetko","year":"1996","journal-title":"J. Chem. Inf. Comput. Sci"},{"key":"2023020210104500100_B58","doi-asserted-by":"crossref","first-page":"8348","DOI":"10.1073\/pnas.0832373100","article-title":"A Bayesian framework for combining heterogeneous data sources for gene function prediction (in Saccharomyces cerevisiae)","volume":"100","author":"Troyanskaya","year":"2003","journal-title":"Proc. Natl Acad. Sci. USA"},{"key":"2023020210104500100_B59","doi-asserted-by":"crossref","first-page":"267","DOI":"10.1016\/j.sbi.2005.05.010","article-title":"Automatic annotation of protein function","volume":"15","author":"Valencia","year":"2005","journal-title":"Curr. Opin. Struct. Biol"},{"key":"2023020210104500100_B60","doi-asserted-by":"crossref","first-page":"697","DOI":"10.1038\/nbt825","article-title":"Global protein function prediction from protein-protein interaction networks","volume":"21","author":"Vazquez","year":"2003","journal-title":"Nat. Biotechnol"},{"key":"2023020210104500100_B61","doi-asserted-by":"crossref","first-page":"D358","DOI":"10.1093\/nar\/gkl825","article-title":"STRING 7\u2014recent developments in the integration and prediction of protein interactions","volume":"35","author":"von Mering","year":"2007","journal-title":"Nucleic Acids Res"},{"key":"2023020210104500100_B62","doi-asserted-by":"crossref","first-page":"645","DOI":"10.1016\/S0960-894X(01)81246-4","article-title":"The Use of Neural Networks for Variable Selection in QSAR","volume":"3","author":"Wikel","year":"1993","journal-title":"Bioorg. Med. Chem. Lett"},{"key":"2023020210104500100_B63","doi-asserted-by":"crossref","first-page":"149","DOI":"10.1016\/0097-8485(93)85006-X","article-title":"Statistics of local complexity in amino acid sequences and sequence databases","volume":"17","author":"Wootton","year":"1993","journal-title":"Comput. Chem"},{"key":"2023020210104500100_B64","doi-asserted-by":"crossref","first-page":"49","DOI":"10.1093\/nar\/28.1.49","article-title":"ProtoMap: automatic classification of protein sequences and hierarchy of protein families","volume":"28","author":"Yona","year":"2000","journal-title":"Nucleic Acids Res"},{"key":"2023020210104500100_B65","doi-asserted-by":"crossref","first-page":"2163","DOI":"10.1093\/bioinformatics\/btm291","article-title":"Total ancestry measure: quantifying the similarity in tree-like classification, with genomic applications","volume":"23","author":"Yu","year":"2007","journal-title":"Bioinformatics"}],"container-title":["Bioinformatics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/academic.oup.com\/bioinformatics\/article-pdf\/24\/5\/621\/49050534\/bioinformatics_24_5_621.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/academic.oup.com\/bioinformatics\/article-pdf\/24\/5\/621\/49050534\/bioinformatics_24_5_621.pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2023,2,2]],"date-time":"2023-02-02T11:47:23Z","timestamp":1675338443000},"score":1,"resource":{"primary":{"URL":"https:\/\/academic.oup.com\/bioinformatics\/article\/24\/5\/621\/201143"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2008,1,3]]},"references-count":65,"journal-issue":{"issue":"5","published-print":{"date-parts":[[2008,3,1]]}},"URL":"https:\/\/doi.org\/10.1093\/bioinformatics\/btm633","relation":{},"ISSN":["1367-4811","1367-4803"],"issn-type":[{"value":"1367-4811","type":"electronic"},{"value":"1367-4803","type":"print"}],"subject":[],"published-other":{"date-parts":[[2008,3,1]]},"published":{"date-parts":[[2008,1,3]]}}}