{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,2,19]],"date-time":"2026-02-19T03:55:11Z","timestamp":1771473311554,"version":"3.50.1"},"reference-count":63,"publisher":"Springer Science and Business Media LLC","issue":"1","content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["BMC Bioinformatics"],"published-print":{"date-parts":[[2008,12]]},"abstract":"<jats:title>Abstract<\/jats:title>\n          <jats:sec>\n            <jats:title>Background<\/jats:title>\n            <jats:p>Residue depth allows determining how deeply a given residue is buried, in contrast to the solvent accessibility that differentiates between buried and solvent-exposed residues. When compared with the solvent accessibility, the depth allows studying deep-level structures and functional sites, and formation of the protein folding nucleus. Accurate prediction of residue depth would provide valuable information for fold recognition, prediction of functional sites, and protein design.<\/jats:p>\n          <\/jats:sec>\n          <jats:sec>\n            <jats:title>Results<\/jats:title>\n            <jats:p>A new method, RDPred, for the real-value depth prediction from protein sequence is proposed. RDPred combines information extracted from the sequence, PSI-BLAST scoring matrices, and secondary structure predicted with PSIPRED. Three-fold\/ten-fold cross validation based tests performed on three independent, low-identity datasets show that the distance based depth (computed using MSMS) predicted by RDPred is characterized by 0.67\/0.67, 0.66\/0.67, and 0.64\/0.65 correlation with the actual depth, by the mean absolute errors equal 0.56\/0.56, 0.61\/0.60, and 0.58\/0.57, and by the mean relative errors equal 17.0%\/16.9%, 18.2%\/18.1%, and 17.7%\/17.6%, respectively. The mean absolute and the mean relative errors are shown to be statistically significantly better when compared with a method recently proposed by Yuan and Wang [Proteins 2008; 70:509\u2013516]. The results show that three-fold cross validation underestimates the variability of the prediction quality when compared with the results based on the ten-fold cross validation. We also show that the hydrophilic and flexible residues are predicted more accurately than hydrophobic and rigid residues. Similarly, the charged residues that include Lys, Glu, Asp, and Arg are the most accurately predicted. Our analysis reveals that evolutionary information encoded using PSSM is characterized by stronger correlation with the depth for hydrophilic amino acids (AAs) and aliphatic AAs when compared with hydrophobic AAs and aromatic AAs. Finally, we show that the secondary structure of coils and strands is useful in depth prediction, in contrast to helices that have relatively uniform distribution over the protein depth. Application of the predicted residue depth to prediction of buried\/exposed residues shows consistent improvements in detection rates of both buried and exposed residues when compared with the competing method. Finally, we contrasted the prediction performance among distance based (MSMS and DPX) and volume based (SADIC) depth definitions. We found that the distance based indices are harder to predict due to the more complex nature of the corresponding depth profiles.<\/jats:p>\n          <\/jats:sec>\n          <jats:sec>\n            <jats:title>Conclusion<\/jats:title>\n            <jats:p>The proposed method, RDPred, provides statistically significantly better predictions of residue depth when compared with the competing method. The predicted depth can be used to provide improved prediction of both buried and exposed residues. The prediction of exposed residues has implications in characterization\/prediction of interactions with ligands and other proteins, while the prediction of buried residues could be used in the context of folding predictions and simulations.<\/jats:p>\n          <\/jats:sec>","DOI":"10.1186\/1471-2105-9-388","type":"journal-article","created":{"date-parts":[[2008,9,20]],"date-time":"2008-09-20T18:13:30Z","timestamp":1221934410000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":30,"title":["Sequence based residue depth prediction using evolutionary information and predicted secondary structure"],"prefix":"10.1186","volume":"9","author":[{"given":"Hua","family":"Zhang","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Tuo","family":"Zhang","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Ke","family":"Chen","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Shiyi","family":"Shen","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jishou","family":"Ruan","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Lukasz","family":"Kurgan","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2008,9,20]]},"reference":[{"key":"2373_CR1","doi-asserted-by":"publisher","first-page":"223","DOI":"10.1126\/science.181.4096.223","volume":"181","author":"CB Anfinsen","year":"1973","unstructured":"Anfinsen CB: Principles that govern the folding of protein chains. Science 1973, 181: 223\u2013230. 10.1126\/science.181.4096.223","journal-title":"Science"},{"issue":"Suppl 6","key":"2373_CR2","doi-asserted-by":"publisher","first-page":"457","DOI":"10.1002\/prot.10552","volume":"53","author":"P Bradley","year":"2003","unstructured":"Bradley P, Chivian D, Meiler J, Misura K, Rohl C, Schief W, Wedemeyer W, Schueler-Furman O, Murphy P, Schonbrun J, Strauss C, Baker D: Rosetta predictions in CASP5: Successes, failures, and prospects for complete automation. Proteins 2003, 53(Suppl 6):457\u2013468. 10.1002\/prot.10552","journal-title":"Proteins"},{"issue":"Suppl 8","key":"2373_CR3","doi-asserted-by":"publisher","first-page":"3","DOI":"10.1002\/prot.21767","volume":"69","author":"J Moult","year":"2007","unstructured":"Moult J, Fidelis K, Kryshtafovych A, Rost B, Hubbard T, Tramontano A: Critical assessment of methods of protein structure prediction \u2013 Round VII. Proteins 2007, 69(Suppl 8):3\u20139. 10.1002\/prot.21767","journal-title":"Proteins"},{"key":"2373_CR4","doi-asserted-by":"publisher","first-page":"379","DOI":"10.1016\/0022-2836(71)90324-X","volume":"55","author":"B Lee","year":"1971","unstructured":"Lee B, Richards F: The interpretation of protein structures: estimation of static accessibility. J Mol Biol 1971, 55: 379\u2013400. 10.1016\/0022-2836(71)90324-X","journal-title":"J Mol Biol"},{"key":"2373_CR5","doi-asserted-by":"publisher","first-page":"709","DOI":"10.1126\/science.6879170","volume":"221","author":"ML Connoly","year":"1983","unstructured":"Connoly ML: Solvent accessibility surfaces of protein and nucleic acids. Science 1983, 221: 709\u2013713. 10.1126\/science.6879170","journal-title":"Science"},{"key":"2373_CR6","doi-asserted-by":"publisher","first-page":"199","DOI":"10.1038\/319199a0","volume":"319","author":"D Eisenberg","year":"1986","unstructured":"Eisenberg D, McLachlan AD: Solvation energy in protein folding and binding. Nature 1986, 319: 199\u2013203. 10.1038\/319199a0","journal-title":"Nature"},{"key":"2373_CR7","doi-asserted-by":"publisher","first-page":"549","DOI":"10.1093\/protein\/12.7.549","volume":"12","author":"MM Gromiha","year":"1999","unstructured":"Gromiha MM, Oobatake M, Kono H, Uedaira H, Sarai A: Role of structural and sequence information in the prediction of protein stability changes, comparison between buried and partially buried mutations. Protein Engineering 1999, 12: 549\u2013555. 10.1093\/protein\/12.7.549","journal-title":"Protein Engineering"},{"issue":"12","key":"2373_CR8","doi-asserted-by":"publisher","first-page":"1456","DOI":"10.1093\/bioinformatics\/btl102","volume":"22","author":"J Cheng","year":"2006","unstructured":"Cheng J, Baldi P: A machine learning information retrieval approach to protein fold recognition. Bioinformatics 2006, 22(12):1456\u201363. 10.1093\/bioinformatics\/btl102","journal-title":"Bioinformatics"},{"key":"2373_CR9","doi-asserted-by":"publisher","first-page":"636","DOI":"10.1002\/prot.21459","volume":"68","author":"S Liu","year":"2007","unstructured":"Liu S, Zhang C, Liang S, Zhou Y: Fold recognition by concurrent use of solvent accessibility and residue depth. Proteins 2007, 68: 636\u2013645. 10.1002\/prot.21459","journal-title":"Proteins"},{"key":"2373_CR10","doi-asserted-by":"publisher","first-page":"216","DOI":"10.1002\/prot.340200303","volume":"20","author":"B Rost","year":"1994","unstructured":"Rost B, Sander C: Conservation and prediction of solvent accessibility in protein families. Proteins 1994, 20: 216\u2013226. 10.1002\/prot.340200303","journal-title":"Proteins"},{"key":"2373_CR11","doi-asserted-by":"publisher","first-page":"629","DOI":"10.1002\/prot.10328","volume":"50","author":"S Ahmad","year":"2003","unstructured":"Ahmad S, Gromiha MM, Sarai A: Real value prediction of solvent accessibility from amino acid sequence. Proteins 2003, 50: 629\u2013635. 10.1002\/prot.10328","journal-title":"Proteins"},{"key":"2373_CR12","doi-asserted-by":"publisher","first-page":"558","DOI":"10.1002\/prot.20234","volume":"57","author":"Z Yuan","year":"2004","unstructured":"Yuan Z, Huang B: Prediction of protein accessible surface areas by support vector regression. Proteins 2004, 57: 558\u2013564. 10.1002\/prot.20234","journal-title":"Proteins"},{"issue":"2","key":"2373_CR13","doi-asserted-by":"publisher","first-page":"318","DOI":"10.1002\/prot.20630","volume":"61","author":"A Garg","year":"2005","unstructured":"Garg A, Kaur H, Raghava GP: Real value prediction of solvent accessibility in proteins using multiple sequence alignment and secondary structure. Proteins 2005, 61(2):318\u201324. 10.1002\/prot.20630","journal-title":"Proteins"},{"key":"2373_CR14","doi-asserted-by":"publisher","first-page":"481","DOI":"10.1002\/prot.20620","volume":"61","author":"JY Wang","year":"2005","unstructured":"Wang JY, Lee HM, Ahmad S: Prediction and evolutionary information analysis of protein solvent accessibility using multiple linear regression. Proteins 2005, 61: 481\u2013491. 10.1002\/prot.20620","journal-title":"Proteins"},{"key":"2373_CR15","doi-asserted-by":"publisher","first-page":"542","DOI":"10.1002\/prot.20883","volume":"63","author":"MN Nguyen","year":"2006","unstructured":"Nguyen MN, Rajapakse JC: Two-stage support vector regression approach for predicting accessible surface areas of amino acids. Proteins 2006, 63: 542\u2013550. 10.1002\/prot.20883","journal-title":"Proteins"},{"key":"2373_CR16","doi-asserted-by":"publisher","first-page":"1063","DOI":"10.1021\/pr050397b","volume":"5","author":"Z Yuan","year":"2006","unstructured":"Yuan Z, Zhang F, Davis MJ, Boden M, Teasdale RD: Predicting the solvent accessibility of transmembrane residues from protein sequence. J Proteome Res 2006, 5: 1063\u20131070. 10.1021\/pr050397b","journal-title":"J Proteome Res"},{"key":"2373_CR17","doi-asserted-by":"publisher","first-page":"82","DOI":"10.1002\/prot.21422","volume":"68","author":"JY Wang","year":"2007","unstructured":"Wang JY, Lee HM, Ahmad S: SVM-Cabins: Prediction of Solvent Accessibility Using Accumulation Cutoff Set and Support Vector Machine. Proteins 2007, 68: 82\u201391. 10.1002\/prot.21422","journal-title":"Proteins"},{"key":"2373_CR18","doi-asserted-by":"publisher","first-page":"85","DOI":"10.1016\/S0006-3495(04)74086-2","volume":"86","author":"AR Atilgan","year":"2004","unstructured":"Atilgan AR, Akan P, Baysal C: Small-World Communication of Residues and Significance for Protein Dynamics. Biophys J 2004, 86: 85\u201391.","journal-title":"Biophys J"},{"key":"2373_CR19","doi-asserted-by":"publisher","first-page":"6388","DOI":"10.1073\/pnas.87.16.6388","volume":"87","author":"HS Chan","year":"1990","unstructured":"Chan HS, Dill KA: Origins of structure in globular proteins. Proc Natl Acad Sci USA 1990, 87: 6388\u20136392. 10.1073\/pnas.87.16.6388","journal-title":"Proc Natl Acad Sci USA"},{"key":"2373_CR20","doi-asserted-by":"publisher","first-page":"105","DOI":"10.1016\/S0022-2836(02)01036-7","volume":"324","author":"GJ Bartlett","year":"2002","unstructured":"Bartlett GJ, Porter CT, Borkakoti N, Thornton JM: Analysis of Catalytic Residues in Enzyme Active Sites. J Mol Bio 2002, 324: 105\u2013121. 10.1016\/S0022-2836(02)01036-7","journal-title":"J Mol Bio"},{"key":"2373_CR21","doi-asserted-by":"publisher","first-page":"413","DOI":"10.1016\/0022-2836(91)90722-I","volume":"218","author":"TG Pedersen","year":"1991","unstructured":"Pedersen TG, Sigurskjold BW, Andersen KV, Kjaer M, Poulsen FM, Dobson CM, Redfield C: A nuclear-magnetic-resonance study of the hydrogen-exchange behavior of lysozyme in crystals and solution. J Mol Biol 1991, 218: 413\u2013426. 10.1016\/0022-2836(91)90722-I","journal-title":"J Mol Biol"},{"key":"2373_CR22","doi-asserted-by":"publisher","first-page":"723","DOI":"10.1016\/S0969-2126(99)80097-5","volume":"7","author":"S Chakravarty","year":"1999","unstructured":"Chakravarty S, Varadarajan R: Residue depth: a novel parameter for the analysis of protein structure and stability. Structure 1999, 7: 723\u2013732. 10.1016\/S0969-2126(99)80097-5","journal-title":"Structure"},{"key":"2373_CR23","doi-asserted-by":"publisher","first-page":"2553","DOI":"10.1016\/S0006-3495(03)75060-7","volume":"84","author":"A Pintar","year":"2003","unstructured":"Pintar A, Carugo O, Pongor S: Atom depth as a descriptor of the protein interior. Biophys J 2003, 84: 2553\u20132561.","journal-title":"Biophys J"},{"key":"2373_CR24","doi-asserted-by":"publisher","first-page":"313","DOI":"10.1093\/bioinformatics\/19.2.313","volume":"19","author":"A Pintar","year":"2003","unstructured":"Pintar A, Carugo O, Pongor S: DPX, for the analysis of the protein core. Bioinformatics 2003, 19: 313\u2013314. 10.1093\/bioinformatics\/19.2.313","journal-title":"Bioinformatics"},{"issue":"12","key":"2373_CR25","doi-asserted-by":"publisher","first-page":"2856","DOI":"10.1093\/bioinformatics\/bti444","volume":"21","author":"D Varrazzo","year":"2005","unstructured":"Varrazzo D, Bernini A, Spiga O, Ciutti A, Chiellini SV, Bracci L, Niccolai N: Three-dimensional computation of atom depth in complex molecular structures. Bioinformatics 2005, 21(12):2856\u20132860. 10.1093\/bioinformatics\/bti444","journal-title":"Bioinformatics"},{"key":"2373_CR26","doi-asserted-by":"publisher","first-page":"719","DOI":"10.1016\/S0022-2836(03)00515-1","volume":"330","author":"A Gutteridge","year":"2003","unstructured":"Gutteridge A, Bartlett GJ, Thornton JM: Using a neural network and spatial clustering to predict the location of active sites in enzymes. J Mol Biol 2003, 330: 719\u2013734. 10.1016\/S0022-2836(03)00515-1","journal-title":"J Mol Biol"},{"key":"2373_CR27","doi-asserted-by":"publisher","first-page":"19","DOI":"10.1186\/1472-6807-8-19","volume":"8","author":"J Kitchen","year":"2008","unstructured":"Kitchen J, Saunders RE, Warwicker J: Charge environments around phosphorylation sites in proteins. BMC Struct Biol 2008, 8: 19. 10.1186\/1472-6807-8-19","journal-title":"BMC Struct Biol"},{"key":"2373_CR28","doi-asserted-by":"publisher","first-page":"1005","DOI":"10.1002\/prot.20007","volume":"55","author":"H Zhou","year":"2004","unstructured":"Zhou H, Zhou Y: Single-body residue-level knowledge-based energy score combined with sequence-profile and secondary structure information for fold recognition. Proteins 2004, 55: 1005\u20131013. 10.1002\/prot.20007","journal-title":"Proteins"},{"key":"2373_CR29","doi-asserted-by":"publisher","first-page":"584","DOI":"10.1002\/prot.20529","volume":"60","author":"A Pintar","year":"2005","unstructured":"Pintar A, Pongor S: The \"first in-last out\" hypothesis on protein folding revisited. Proteins 2005, 60: 584\u2013590. 10.1002\/prot.20529","journal-title":"Proteins"},{"key":"2373_CR30","doi-asserted-by":"publisher","first-page":"509","DOI":"10.1002\/prot.21545","volume":"70","author":"Z Yuan","year":"2008","unstructured":"Yuan Z, Wang ZX: Quantifying the relationship of protein burying depth and sequence. Proteins 2008, 70: 509\u2013516. 10.1002\/prot.21545","journal-title":"Proteins"},{"key":"2373_CR31","doi-asserted-by":"publisher","first-page":"199","DOI":"10.1023\/B:STCO.0000035301.49549.88","volume":"14","author":"AJ Smola","year":"2004","unstructured":"Smola AJ, Sch\u00f6lkopf B: A tutorial on support vector regression. Statistics and Computing 2004, 14: 199\u2013222. 10.1023\/B:STCO.0000035301.49549.88","journal-title":"Statistics and Computing"},{"key":"2373_CR32","doi-asserted-by":"publisher","first-page":"248","DOI":"10.1186\/1471-2105-6-248","volume":"6","author":"Z Yuan","year":"2005","unstructured":"Yuan Z: Better prediction of protein contact number using a support vector regression analysis of amino acid sequence. BMC Bioinformatics 2005, 6: 248. 10.1186\/1471-2105-6-248","journal-title":"BMC Bioinformatics"},{"key":"2373_CR33","doi-asserted-by":"publisher","first-page":"59","DOI":"10.1186\/1471-2105-6-59","volume":"6","author":"GP Raghava","year":"2005","unstructured":"Raghava GP, Han JH: Correlation and prediction of gene expression level from amino acid and dipeptide composition of its protein. BMC Bioinformatics 2005, 6: 59. 10.1186\/1471-2105-6-59","journal-title":"BMC Bioinformatics"},{"key":"2373_CR34","doi-asserted-by":"publisher","first-page":"425","DOI":"10.1186\/1471-2105-7-425","volume":"7","author":"J Song","year":"2006","unstructured":"Song J, Burrage K: Predicting residue-wise contact orders in proteins by support vector regression. BMC Bioinformatics 2006, 7: 425. 10.1186\/1471-2105-7-425","journal-title":"BMC Bioinformatics"},{"key":"2373_CR35","doi-asserted-by":"publisher","first-page":"182","DOI":"10.1186\/1471-2105-7-182","volume":"7","author":"W Liu","year":"2006","unstructured":"Liu W, Meng X, Xu Q, Flower DR, Li T: Quantitative prediction of mouse class I MHC peptide binding affinity using support vector machine regression (SVR) models. BMC Bioinformatics 2006, 7: 182. 10.1186\/1471-2105-7-182","journal-title":"BMC Bioinformatics"},{"key":"2373_CR36","doi-asserted-by":"publisher","first-page":"3389","DOI":"10.1093\/nar\/25.17.3389","volume":"25","author":"SF Altschul","year":"1997","unstructured":"Altschul SF, Madden TL, Sch\u00e4ffer AA, Zhang J, Zhang Z, Miller W, Lipman DJ: Gapped BLAST and PSI-BLAST: a new generation of protein database search programs. Nucleic Acids Res 1997, 25: 3389\u20133402. 10.1093\/nar\/25.17.3389","journal-title":"Nucleic Acids Res"},{"key":"2373_CR37","doi-asserted-by":"publisher","first-page":"195","DOI":"10.1006\/jmbi.1999.3091","volume":"292","author":"DT Jones","year":"1999","unstructured":"Jones DT: Protein secondary structure prediction based on position-specific scoring matrices. J Mol Biol 1999, 292: 195\u2013202. 10.1006\/jmbi.1999.3091","journal-title":"J Mol Biol"},{"key":"2373_CR38","doi-asserted-by":"publisher","first-page":"W36","DOI":"10.1093\/nar\/gki410","volume-title":"Nucl Acids Res","author":"K Bryson","year":"2005","unstructured":"Bryson K, McGuffin LJ, Marsden RL, Ward JJ, Sodhi JS, Jones DT: Protein structure prediction servers at University College London. Nucl Acids Res 2005, (33 Web Server):W36\u201338. 10.1093\/nar\/gki410"},{"key":"2373_CR39","doi-asserted-by":"publisher","first-page":"492","DOI":"10.1093\/nar\/gkg022","volume":"31","author":"T Noguchi","year":"2003","unstructured":"Noguchi T, Akiyama Y: PDB-REPRDB: a database of representative protein chains from the Protein Data Bank (PDB) in 2003. Nucleic Acids Res 2003, 31: 492\u2013493. 10.1093\/nar\/gkg022","journal-title":"Nucleic Acids Res"},{"key":"2373_CR40","doi-asserted-by":"publisher","first-page":"235","DOI":"10.1093\/nar\/28.1.235","volume":"28","author":"HM Berman","year":"2000","unstructured":"Berman HM, Westbrook J, Feng Z, Gilliland G, Bhat TN, Weissig H, Shindyalov IN, Bourne PE: The Protein Data Bank. Nucleic Acids Res 2000, 28: 235\u2013242. 10.1093\/nar\/28.1.235","journal-title":"Nucleic Acids Res"},{"key":"2373_CR41","doi-asserted-by":"publisher","first-page":"1658","DOI":"10.1093\/bioinformatics\/btl158","volume":"22","author":"W Li","year":"2006","unstructured":"Li W, Godzik A: Cd-hit: a fast program for clustering and comparing large sets of protein or nucleotide sequences. Bioinformatics 2006, 22: 1658\u20139. 10.1093\/bioinformatics\/btl158","journal-title":"Bioinformatics"},{"key":"2373_CR42","doi-asserted-by":"publisher","first-page":"443","DOI":"10.1016\/0022-2836(70)90057-4","volume":"48","author":"SB Needleman","year":"1970","unstructured":"Needleman SB, Wunsch CD: A general method applicable to the search for similarities in the amino acid sequence of two proteins. J Mol Biol 1970, 48: 443\u2013453. 10.1016\/0022-2836(70)90057-4","journal-title":"J Mol Biol"},{"key":"2373_CR43","doi-asserted-by":"publisher","first-page":"305","DOI":"10.1002\/(SICI)1097-0282(199603)38:3<305::AID-BIP4>3.0.CO;2-Y","volume":"38","author":"MF Sanner","year":"1996","unstructured":"Sanner MF, Olson AJ, Spehner JC: Reduced surface: an efficient way to compute molecular surfaces. Biopolymers 1996, 38: 305\u2013320. Publisher Full Text 10.1002\/(SICI)1097-0282(199603)38:3<305::AID-BIP4>3.0.CO;2-Y","journal-title":"Biopolymers"},{"key":"2373_CR44","doi-asserted-by":"publisher","first-page":"38","DOI":"10.1002\/prot.20379","volume":"59","author":"T Hamelryck","year":"2005","unstructured":"Hamelryck T: An amino acid has two sides: a new 2D measure provides a different view of solvent exposure. Proteins 2005, 59: 38\u201348. 10.1002\/prot.20379","journal-title":"Proteins"},{"key":"2373_CR45","volume-title":"NACCESS","author":"SJ Hubbard","year":"1993","unstructured":"Hubbard SJ, Thornton JM: NACCESS. Department of Biochemistry and Molecular Biology, University College, London; 1993."},{"key":"2373_CR46","doi-asserted-by":"publisher","first-page":"575","DOI":"10.1002\/prot.21036","volume":"64","author":"G Karypis","year":"2006","unstructured":"Karypis G: YASSPP: Better Kernels and Coding Schemes Lead to Improvements in Protein Secondary Structure Prediction. Proteins 2006, 64: 575\u2013586. 10.1002\/prot.21036","journal-title":"Proteins"},{"key":"2373_CR47","doi-asserted-by":"publisher","first-page":"2628","DOI":"10.1093\/bioinformatics\/btl453","volume":"22","author":"F Birzele","year":"2006","unstructured":"Birzele F, Kramer S: A new representation for protein secondary structure prediction based on frequent patterns. Bioinformatics 2006, 22: 2628\u201334. 10.1093\/bioinformatics\/btl453","journal-title":"Bioinformatics"},{"key":"2373_CR48","doi-asserted-by":"publisher","first-page":"2843","DOI":"10.1093\/bioinformatics\/btm475","volume":"23","author":"K Chen","year":"2007","unstructured":"Chen K, Kurgan L: PFRES: Protein Fold Classification by Using Evolutionary Information and Predicted Secondary Structure. Bioinformatics 2007, 23: 2843\u20132850. 10.1093\/bioinformatics\/btm475","journal-title":"Bioinformatics"},{"key":"2373_CR49","doi-asserted-by":"publisher","first-page":"8942","DOI":"10.1073\/pnas.0402659101","volume":"101","author":"DN Ivankov","year":"2004","unstructured":"Ivankov DN, Finkelstein AV: Prediction of protein folding rates from the amino acid sequence-predicted secondary structure. Proc Nat Acad Sci USA 2004, 101: 8942\u20134. 10.1073\/pnas.0402659101","journal-title":"Proc Nat Acad Sci USA"},{"key":"2373_CR50","doi-asserted-by":"publisher","first-page":"828","DOI":"10.1002\/prot.20461","volume":"59","author":"PF Fuchs","year":"2005","unstructured":"Fuchs PF, Alix AJ: High accuracy prediction of beta-turns and their types using propensities and multiple alignments. Proteins 2005, 59: 828\u2013839. 10.1002\/prot.20461","journal-title":"Proteins"},{"key":"2373_CR51","doi-asserted-by":"publisher","first-page":"49","DOI":"10.1002\/prot.21062","volume":"65","author":"Y Wang","year":"2006","unstructured":"Wang Y, Xue Z, Xu J: Better prediction of the location of alpha-turns in proteins with support vector machine. Proteins 2006, 65: 49\u201354. 10.1002\/prot.21062","journal-title":"Proteins"},{"key":"2373_CR52","doi-asserted-by":"publisher","first-page":"175","DOI":"10.1016\/S0969-2126(02)00700-1","volume":"10","author":"CAF Andersen","year":"2002","unstructured":"Andersen CAF, Palmer AG, Brunak S, Rost B: Continuum Secondary Structure Captures Protein Flexibility. Structure 2002, 10: 175\u2013184. 10.1016\/S0969-2126(02)00700-1","journal-title":"Structure"},{"key":"2373_CR53","volume-title":"Statistical learning theory","author":"V Vapnik","year":"1998","unstructured":"Vapnik V: Statistical learning theory. New York: Wiley; 1998."},{"key":"2373_CR54","doi-asserted-by":"publisher","first-page":"905","DOI":"10.1002\/prot.20375","volume":"58","author":"Z Yuan","year":"2005","unstructured":"Yuan Z, Bailey TL, Teasdale RD: Prediction of protein B-factor profiles. Proteins 2005, 58: 905\u2013912. 10.1002\/prot.20375","journal-title":"Proteins"},{"key":"2373_CR55","doi-asserted-by":"publisher","first-page":"996","DOI":"10.1136\/bmj.309.6960.996","volume":"309","author":"DG Altman","year":"1994","unstructured":"Altman DG, Bland JM: Quartiles, quintiles, centiles, and other quantiles. BMJ 1994, 309: 996.","journal-title":"BMJ"},{"key":"2373_CR56","first-page":"856","volume-title":"Proceedings of the 10th International Conference on Machine Learning","author":"L Yu","year":"2003","unstructured":"Yu L, Liu H: Feature Selection for High-Dimensional Data: A Fast Correlation-Based Filter Solution. Proceedings of the 10th International Conference on Machine Learning 2003, 856\u2013863."},{"key":"2373_CR57","doi-asserted-by":"publisher","first-page":"415","DOI":"10.1109\/TNN.2002.1000139","volume":"13","author":"CW Hsu","year":"2002","unstructured":"Hsu CW, Lin CJ: A comparison on methods for multi-class support vector machines. IEEE Trans Neural Networks 2002, 13: 415\u2013425. 10.1109\/72.991427","journal-title":"IEEE Trans Neural Networks"},{"key":"2373_CR58","doi-asserted-by":"publisher","first-page":"2577","DOI":"10.1002\/bip.360221211","volume":"22","author":"W Kabsch","year":"1983","unstructured":"Kabsch W, Sander C: Dictionary of protein secondary structure: pattern recognition of hydrogen-bonded and geometrical features. Biopolymers 1983, 22: 2577\u20132637. 10.1002\/bip.360221211","journal-title":"Biopolymers"},{"key":"2373_CR59","doi-asserted-by":"publisher","first-page":"374","DOI":"10.1093\/nar\/28.1.374","volume":"28","author":"S Kawashima","year":"2000","unstructured":"Kawashima S, Kanehisa M: AAindex: amino acid index database. Nucleic Acids Res 2000, 28: 374. 10.1093\/nar\/28.1.374","journal-title":"Nucleic Acids Res"},{"key":"2373_CR60","doi-asserted-by":"publisher","first-page":"479","DOI":"10.1016\/0022-2836(83)90041-4","volume":"171","author":"RM Sweet","year":"1983","unstructured":"Sweet RM, Eisenberg D: Correlation of sequence hydrophobicities measures similarity in three dimensional protein structure. J Mol Biol 1983, 171: 479\u2013488. 10.1016\/0022-2836(83)90041-4","journal-title":"J Mol Biol"},{"key":"2373_CR61","doi-asserted-by":"publisher","first-page":"141","DOI":"10.1002\/prot.340190207","volume":"19","author":"M Vihinen","year":"1994","unstructured":"Vihinen M, Torkkila E, Riikonen P: Accuracy of protein flexibility predictions. Proteins 1994, 19: 141\u2013149. 10.1002\/prot.340190207","journal-title":"Proteins"},{"key":"2373_CR62","first-page":"366","volume-title":"Proceedings of the 2006 IEEE Symposium on Computational Intelligence in Bioinformatics and Computational Biology Toronto, Ontario, Canada","author":"K Chen","year":"2006","unstructured":"Chen K, Kurgan LA, Ruan J: Optimization of the Sliding Window Size for Protein Structure Prediction. Proceedings of the 2006 IEEE Symposium on Computational Intelligence in Bioinformatics and Computational Biology Toronto, Ontario, Canada 2006, 366\u2013372."},{"issue":"3","key":"2373_CR63","doi-asserted-by":"publisher","first-page":"198","DOI":"10.1093\/bib\/bbm064","volume":"9","author":"P Sonego","year":"2008","unstructured":"Sonego P, Kocsor A, Pongor S: ROC analysis: applications to the classification of biological sequences and 3D structures. Briefings in Bioinformatics 2008, 9(3):198\u2013209. 10.1093\/bib\/bbm064","journal-title":"Briefings in Bioinformatics"}],"container-title":["BMC Bioinformatics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1186\/1471-2105-9-388.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2021,9,1]],"date-time":"2021-09-01T03:15:55Z","timestamp":1630466155000},"score":1,"resource":{"primary":{"URL":"https:\/\/bmcbioinformatics.biomedcentral.com\/articles\/10.1186\/1471-2105-9-388"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2008,9,20]]},"references-count":63,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2008,12]]}},"alternative-id":["2373"],"URL":"https:\/\/doi.org\/10.1186\/1471-2105-9-388","relation":{},"ISSN":["1471-2105"],"issn-type":[{"value":"1471-2105","type":"electronic"}],"subject":[],"published":{"date-parts":[[2008,9,20]]},"assertion":[{"value":"10 April 2008","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"20 September 2008","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"20 September 2008","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}}],"article-number":"388"}}