{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,2,17]],"date-time":"2026-02-17T14:37:55Z","timestamp":1771339075148,"version":"3.50.1"},"reference-count":23,"publisher":"Oxford University Press (OUP)","issue":"18","content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2005,9,15]]},"abstract":"<jats:title>Abstract<\/jats:title>\n               <jats:p>Motivation: Computational gene finding systems play an important role in finding new human genes, although no systems are yet accurate enough to predict all or even most protein-coding regions perfectly. Ab initio programs can be augmented by evidence such as expression data or protein sequence homology, which improves their performance. The amount of such evidence continues to grow, but computational methods continue to have difficulty predicting genes when the evidence is conflicting or incomplete. Genome annotation pipelines collect a variety of types of evidence about gene structure and synthesize the results, which can then be refined further through manual, expert curation of gene models.<\/jats:p>\n               <jats:p>Results: JIGSAW is a new gene finding system designed to automate the process of predicting gene structure from multiple sources of evidence, with results that often match the performance of human curators. JIGSAW computes the relative weight of different lines of evidence using statistics generated from a training set, and then combines the evidence using dynamic programming. Our results show that JIGSAW's performance is superior to ab initio gene finding methods and to other pipelines such as Ensembl. Even without evidence from alignment to known genes, JIGSAW can substantially improve gene prediction accuracy as compared with existing methods.<\/jats:p>\n               <jats:p>Availability: JIGSAW is available as an open source software package at http:\/\/cbcb.umd.edu\/software\/jigsaw<\/jats:p>\n               <jats:p>Contact: \u00a0jeallen@umiacs.umd.edu<\/jats:p>","DOI":"10.1093\/bioinformatics\/bti609","type":"journal-article","created":{"date-parts":[[2005,8,3]],"date-time":"2005-08-03T02:43:46Z","timestamp":1123037026000},"page":"3596-3603","source":"Crossref","is-referenced-by-count":122,"title":["JIGSAW: integration of multiple sources of evidence for gene prediction"],"prefix":"10.1093","volume":"21","author":[{"given":"Jonathan E.","family":"Allen","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Steven L.","family":"Salzberg","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"286","published-online":{"date-parts":[[2005,8,2]]},"reference":[{"key":"2023060912072553700_B1","doi-asserted-by":"crossref","unstructured":"Allen, J.E., Pertea, M., Salzberg, S.L. 2004Computational gene prediction using multiple sources of evidence. Genome Research14","DOI":"10.1101\/gr.1562804"},{"key":"2023060912072553700_B2","unstructured":"Ashurst, J.L., et al. 2005The Vertebrate genome annotation ({V}ega) database. Nucleic Acids Res.33459\u2013465"},{"key":"2023060912072553700_B3","unstructured":"Bairoch, A., et al. 2005The universal protein resource ({U}ni{P}rot). Nucleic Acids Res.33154\u2013159"},{"key":"2023060912072553700_B4","unstructured":"Buell, C.R., et al. 2005Sequence, annotation, and analysis of synteny between rice chromosome 3 and diverged grass speices. Genome Res.  in press"},{"key":"2023060912072553700_B5","unstructured":"Burge, C. and Karlin, S. 1997Prediction of complete gene structures in human genomic DNA. J. Mol. Biol.26878\u201384"},{"key":"2023060912072553700_B6","doi-asserted-by":"crossref","unstructured":"Curwen, V., et al. 2004The Ensembl automatic gene annotation system. Genome Res.14942\u2013950","DOI":"10.1101\/gr.1858004"},{"key":"2023060912072553700_B7","unstructured":"EGASP. 2005Gene prediction workshop.   http:\/\/genome.imim.es\/gencode\/workshop2005.html"},{"key":"2023060912072553700_B8","doi-asserted-by":"crossref","unstructured":"Flicek, P., et al. 2003Leveraging the mouse genome for gene prediction in human: from whole-genome shotgun reads to a global synteny map. Genome Res.1346\u201354","DOI":"10.1101\/gr.830003"},{"key":"2023060912072553700_B9","unstructured":"Guigo, R. 1998Assembling genes from predicted exons in linear time with dynamic programming. J. Comput. Biol.5681\u2013702"},{"key":"2023060912072553700_B10","unstructured":"International Human Genome Sequencing Consortium. 2001Initial sequencing and analysis of the human genome. Nature409860\u2013921"},{"key":"2023060912072553700_B11","unstructured":"Kent, W.J. 2002BLAT\u2014the BLAST-like alignment tool. Genome Res.12656\u2013664"},{"key":"2023060912072553700_B12","doi-asserted-by":"crossref","unstructured":"Lee, Y., et al. 2005The TIGR gene indices: clustering and assembling EST and known genes and integration with eukaryotic genomes. Nucleic Acids Res.3371\u201374","DOI":"10.1093\/nar\/gki064"},{"key":"2023060912072553700_B13","unstructured":"Loftus, B.J., et al. 2005The genome of the basidiomycetous yeast and human pathogen Cryptococcus neoformans. Science3071321\u20131324"},{"key":"2023060912072553700_B14","unstructured":"Majoros, W.H., et al. 2004Tigr{S}can and Glimmer{HMM}: two open source ab initio eukaryotic gene-finders. Bioinformatics202878\u20132879"},{"key":"2023060912072553700_B15","doi-asserted-by":"crossref","unstructured":"Murthy, S.K., et al. 1994A system for induction of oblique decision trees. J. Artif. Intell. Res.21\u201332","DOI":"10.1613\/jair.63"},{"key":"2023060912072553700_B16","unstructured":"Parra, G., et al. 2003Comparative gene prediction in human and mouse. Genome Res.13108\u2013117"},{"key":"2023060912072553700_B17","doi-asserted-by":"crossref","unstructured":"Pruitt, K.D., et al. 2005NCBI Reference Sequence ({R}ef{S}eq): a curated non-redundant sequence database of genomes, transcripts and proteins. Nucleic Acids Res.1501\u2013504","DOI":"10.1093\/nar\/gki025"},{"key":"2023060912072553700_B18","unstructured":"Salzberg, S.L., et al. 1999Interpolated markov models for eukaryotic gene finding. Genomics5924\u201331"},{"key":"2023060912072553700_B19","unstructured":"Sarawagi, S. and Cohen, W.W. 2004Semi-markov conditional random fields for information extraction. Proceedings of the Advances in Neural Information Processing Systems, 17 (NIPS 2004)Vancourer, BC, Canada"},{"key":"2023060912072553700_B20","doi-asserted-by":"crossref","unstructured":"Siepel, A. and Haussler, D. 2003Combining phylogenetic and hidden markov models in biosequence analysis. Proceedings of the 7th Annual International Conference on Computational Molecular Biology (RECOMB 2003)Berlin, Germany ,  pp. 277\u2013286","DOI":"10.1145\/640075.640111"},{"key":"2023060912072553700_B21","unstructured":"The ENCODE Project Consortium. 2004The ENCODE (ENCyclopedia of DNA elements) project. Science306636\u2013640"},{"key":"2023060912072553700_B22","unstructured":"Venter, J.C., et al. 2001The sequence of the human genome. Science2911304\u20131351"},{"key":"2023060912072553700_B23","unstructured":"Wheeler, D.L., et al. 2003Database resources of the national center for biotechnology. Nucleic Acids Res.3128\u201333"}],"container-title":["Bioinformatics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/academic.oup.com\/bioinformatics\/article-pdf\/21\/18\/3596\/50554960\/bioinformatics_21_18_3596.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/academic.oup.com\/bioinformatics\/article-pdf\/21\/18\/3596\/50554960\/bioinformatics_21_18_3596.pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2023,6,9]],"date-time":"2023-06-09T12:07:43Z","timestamp":1686312463000},"score":1,"resource":{"primary":{"URL":"https:\/\/academic.oup.com\/bioinformatics\/article\/21\/18\/3596\/202486"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2005,8,2]]},"references-count":23,"journal-issue":{"issue":"18","published-print":{"date-parts":[[2005,9,15]]}},"URL":"https:\/\/doi.org\/10.1093\/bioinformatics\/bti609","relation":{},"ISSN":["1367-4811","1367-4803"],"issn-type":[{"value":"1367-4811","type":"electronic"},{"value":"1367-4803","type":"print"}],"subject":[],"published-other":{"date-parts":[[2005,9]]},"published":{"date-parts":[[2005,8,2]]}}}