{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,24]],"date-time":"2026-04-24T18:22:20Z","timestamp":1777054940453,"version":"3.51.4"},"reference-count":68,"publisher":"Oxford University Press (OUP)","issue":"1","license":[{"start":{"date-parts":[[2021,9,15]],"date-time":"2021-09-15T00:00:00Z","timestamp":1631664000000},"content-version":"vor","delay-in-days":1,"URL":"https:\/\/creativecommons.org\/licenses\/by-nc\/4.0\/"}],"funder":[{"name":"Joint Design of Advanced Computing Solutions for Cancer","award":["JDACS4C"],"award-info":[{"award-number":["JDACS4C"]}]},{"DOI":"10.13039\/100000015","name":"U.S. Department of Energy","doi-asserted-by":"publisher","id":[{"id":"10.13039\/100000015","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100000054","name":"National Cancer Institute","doi-asserted-by":"publisher","id":[{"id":"10.13039\/100000054","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100000002","name":"National Institutes of Health","doi-asserted-by":"publisher","award":["HHSN261200800001E"],"award-info":[{"award-number":["HHSN261200800001E"]}],"id":[{"id":"10.13039\/100000002","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100006224","name":"Argonne National Laboratory","doi-asserted-by":"publisher","award":["DE-AC02-06-CH11357"],"award-info":[{"award-number":["DE-AC02-06-CH11357"]}],"id":[{"id":"10.13039\/100006224","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100006227","name":"Lawrence Livermore National Laboratory","doi-asserted-by":"publisher","award":["DE-AC52-07NA27344"],"award-info":[{"award-number":["DE-AC52-07NA27344"]}],"id":[{"id":"10.13039\/100006227","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100008902","name":"Los Alamos National Laboratory","doi-asserted-by":"publisher","award":["DE-AC5206NA25396"],"award-info":[{"award-number":["DE-AC5206NA25396"]}],"id":[{"id":"10.13039\/100008902","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100006228","name":"Oak Ridge National Laboratory","doi-asserted-by":"publisher","award":["DE-AC05-00OR22725"],"award-info":[{"award-number":["DE-AC05-00OR22725"]}],"id":[{"id":"10.13039\/100006228","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2022,1,17]]},"abstract":"<jats:title>Abstract<\/jats:title><jats:p>To enable personalized cancer treatment, machine learning models have been developed to predict drug response as a function of tumor and drug features. However, most algorithm development efforts have relied on cross-validation within a single study to assess model accuracy. While an essential first step, cross-validation within a biological data set typically provides an overly optimistic estimate of the prediction performance on independent test sets. To provide a more rigorous assessment of model generalizability between different studies, we use machine learning to analyze five publicly available cell line-based data sets: National Cancer Institute 60, ancer Therapeutics Response Portal (CTRP), Genomics of Drug Sensitivity in Cancer, Cancer Cell Line Encyclopedia and Genentech Cell Line Screening Initiative (gCSI). Based on observed experimental variability across studies, we explore estimates of prediction upper bounds. We report performance results of a variety of machine learning models, with a multitasking deep neural network achieving the best cross-study generalizability. By multiple measures, models trained on CTRP yield the most accurate predictions on the remaining testing data, and gCSI is the most predictable among the cell line data sets included in this study. With these experiments and further simulations on partial data, two lessons emerge: (1) differences in viability assays can limit model generalizability across studies and (2) drug diversity, more than tumor diversity, is crucial for raising model generalizability in preclinical screening.<\/jats:p>","DOI":"10.1093\/bib\/bbab356","type":"journal-article","created":{"date-parts":[[2021,9,8]],"date-time":"2021-09-08T11:20:51Z","timestamp":1631100051000},"source":"Crossref","is-referenced-by-count":86,"title":["A cross-study analysis of drug response prediction in cancer cell lines"],"prefix":"10.1093","volume":"23","author":[{"given":"Fangfang","family":"Xia","sequence":"first","affiliation":[{"name":"Argonne National Laboratory"}]},{"given":"Jonathan","family":"Allen","sequence":"additional","affiliation":[{"name":"Lawrence Livermore National Laboratory"}]},{"given":"Prasanna","family":"Balaprakash","sequence":"additional","affiliation":[{"name":"Argonne National Laboratory"}]},{"given":"Thomas","family":"Brettin","sequence":"additional","affiliation":[{"name":"Argonne National Laboratory"}]},{"given":"Cristina","family":"Garcia-Cardona","sequence":"additional","affiliation":[{"name":"Los Alamos National Laboratory"}]},{"given":"Austin","family":"Clyde","sequence":"additional","affiliation":[{"name":"Argonne National Laboratory"},{"name":"University of Chicago"}]},{"given":"Judith","family":"Cohn","sequence":"additional","affiliation":[{"name":"Los Alamos National Laboratory"}]},{"given":"James","family":"Doroshow","sequence":"additional","affiliation":[{"name":"National Cancer Institute"}]},{"given":"Xiaotian","family":"Duan","sequence":"additional","affiliation":[{"name":"University of Chicago"}]},{"given":"Veronika","family":"Dubinkina","sequence":"additional","affiliation":[{"name":"University of Illinois at Urbana-Champaign"}]},{"given":"Yvonne","family":"Evrard","sequence":"additional","affiliation":[{"name":"Frederick National Laboratory for Cancer Research"}]},{"given":"Ya Ju","family":"Fan","sequence":"additional","affiliation":[{"name":"Lawrence Livermore National Laboratory"}]},{"given":"Jason","family":"Gans","sequence":"additional","affiliation":[{"name":"Los Alamos National Laboratory"}]},{"given":"Stewart","family":"He","sequence":"additional","affiliation":[{"name":"Lawrence Livermore National Laboratory"}]},{"given":"Pinyi","family":"Lu","sequence":"additional","affiliation":[{"name":"Frederick National Laboratory for Cancer Research"}]},{"given":"Sergei","family":"Maslov","sequence":"additional","affiliation":[{"name":"University of Illinois at Urbana-Champaign"}]},{"given":"Alexander","family":"Partin","sequence":"additional","affiliation":[{"name":"Argonne National Laboratory"}]},{"given":"Maulik","family":"Shukla","sequence":"additional","affiliation":[{"name":"Argonne National Laboratory"}]},{"given":"Eric","family":"Stahlberg","sequence":"additional","affiliation":[{"name":"Frederick National Laboratory for Cancer Research"}]},{"given":"Justin M","family":"Wozniak","sequence":"additional","affiliation":[{"name":"Argonne National Laboratory"}]},{"given":"Hyunseung","family":"Yoo","sequence":"additional","affiliation":[{"name":"Argonne National Laboratory"}]},{"given":"George","family":"Zaki","sequence":"additional","affiliation":[{"name":"Frederick National Laboratory for Cancer Research"}]},{"given":"Yitan","family":"Zhu","sequence":"additional","affiliation":[{"name":"Argonne National Laboratory"}]},{"given":"Rick","family":"Stevens","sequence":"additional","affiliation":[{"name":"Argonne National Laboratory"},{"name":"University of Chicago"}]}],"member":"286","published-online":{"date-parts":[[2021,9,14]]},"reference":[{"issue":"10","key":"2022011920511078000_ref1","doi-asserted-by":"crossref","first-page":"813","DOI":"10.1038\/nrc1951","article-title":"The NCI60 human tumour cell line anticancer drug screen","volume":"6","author":"Shoemaker","year":"2006","journal-title":"Nat Rev Cancer"},{"issue":"1","key":"2022011920511078000_ref2","doi-asserted-by":"crossref","first-page":"85","DOI":"10.1093\/bioinformatics\/btv529","article-title":"Improved large-scale prediction of growth inhibition patterns using the NCI60 cancer cell line panel","volume":"32","author":"Cort\u00e9s-Ciriano","year":"2016","journal-title":"Bioinformatics"},{"issue":"18","key":"2022011920511078000_ref3","doi-asserted-by":"crossref","first-page":"486","DOI":"10.1186\/s12859-018-2509-3","article-title":"Predicting tumor cell line response to drug pairs with deep learning","volume":"19","author":"Xia","year":"2018","journal-title":"BMC Bioinformatics"},{"issue":"7391","key":"2022011920511078000_ref4","doi-asserted-by":"crossref","first-page":"603","DOI":"10.1038\/nature11003","article-title":"The cancer cell line encyclopedia enables predictive modelling of anticancer drug sensitivity","volume":"483","author":"Barretina","year":"2012","journal-title":"Nature"},{"issue":"7757","key":"2022011920511078000_ref5","doi-asserted-by":"crossref","first-page":"503","DOI":"10.1038\/s41586-019-1186-3","article-title":"Next-generation characterization of the cancer cell line encyclopedia","volume":"569","author":"Ghandi","year":"2019","journal-title":"Nature"},{"issue":"D1","key":"2022011920511078000_ref6","doi-asserted-by":"crossref","first-page":"D955","DOI":"10.1093\/nar\/gks1111","article-title":"Genomics of drug sensitivity in cancer (GDSC): a resource for therapeutic biomarker discovery in cancer cells","volume":"41","author":"Yang","year":"2012","journal-title":"Nucleic Acids Res"},{"issue":"5","key":"2022011920511078000_ref7","doi-asserted-by":"crossref","first-page":"1151","DOI":"10.1016\/j.cell.2013.08.003","article-title":"An interactive resource to identify cancer genetic and lineage dependencies targeted by small molecules","volume":"154","author":"Basu","year":"2013","journal-title":"Cell"},{"issue":"11","key":"2022011920511078000_ref8","doi-asserted-by":"crossref","first-page":"1210","DOI":"10.1158\/2159-8290.CD-15-0235","article-title":"Harnessing connectivity in a large-scale small-molecule sensitivity dataset","volume":"5","author":"Seashore-Ludlow","year":"2015","journal-title":"Cancer Discov"},{"issue":"2","key":"2022011920511078000_ref9","doi-asserted-by":"crossref","first-page":"269","DOI":"10.1158\/1541-7786.MCR-17-0378","article-title":"Precision oncology beyond targeted therapy: combining omics data with machine learning matches the majority of cancer cells to effective therapeutics","volume":"16","author":"Ding","year":"2018","journal-title":"Mol Cancer Res"},{"issue":"19","key":"2022011920511078000_ref10","doi-asserted-by":"crossref","first-page":"3743","DOI":"10.1093\/bioinformatics\/btz158","article-title":"Dr.VAE: improving drug response prediction via modeling of drug perturbation effects","volume":"35","author":"Ramp\u00e1\u0161ek","year":"2019","journal-title":"Bioinformatics"},{"issue":"1","key":"2022011920511078000_ref11","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1038\/s41467-019-09799-2","article-title":"Community assessment to advance computational prediction of cancer drug combinations in a pharmacogenomic screen","volume":"10","author":"Menden","year":"2019","journal-title":"Nat Commun"},{"key":"2022011920511078000_ref12","article-title":"A community challenge for PANcancer drug mechanism of action inference from perturbational profile data","author":"Douglass","year":"2020","journal-title":"bioRxiv"},{"issue":"22","key":"2022011920511078000_ref13","doi-asserted-by":"crossref","first-page":"3907","DOI":"10.1093\/bioinformatics\/bty452","article-title":"Predicting cancer drug response using a recommender system","volume":"34","author":"Suphavilai","year":"2018","journal-title":"Bioinformatics"},{"issue":"1","key":"2022011920511078000_ref14","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1038\/s41467-021-22170-8","article-title":"Drug ranking using machine learning systematically predicts the efficacy of anti-cancer drugs","volume":"12","author":"Gerdes","year":"2021","journal-title":"Nat Commun"},{"issue":"11","key":"2022011920511078000_ref15","doi-asserted-by":"crossref","first-page":"3154","DOI":"10.1109\/JBHI.2020.3004663","article-title":"Q-rank: reinforcement learning for recommending algorithms to predict drug sensitivity to cancer therapy","volume":"24","author":"Daoud","year":"2020","journal-title":"IEEE J Biomed Health Inform"},{"issue":"7","key":"2022011920511078000_ref16","doi-asserted-by":"crossref","first-page":"10883","DOI":"10.18632\/oncotarget.14073","article-title":"The cornucopia of meaningful leads: applying deep adversarial autoencoders for new molecule development in oncology","volume":"8","author":"Kadurin","year":"2017","journal-title":"Oncotarget"},{"issue":"10","key":"2022011920511078000_ref17","doi-asserted-by":"crossref","DOI":"10.1371\/journal.pone.0186906","article-title":"Open source machine-learning algorithms for the prediction of optimal cancer drug therapies","volume":"12","author":"Huang","year":"2017","journal-title":"PLoS One"},{"issue":"1","key":"2022011920511078000_ref18","first-page":"1","article-title":"Comprehensive anticancer drug response prediction based on a simple cell line-drug complex network model","volume":"20","author":"Dong","year":"2019","journal-title":"BMC Bioinformatics"},{"key":"2022011920511078000_ref19","doi-asserted-by":"crossref","first-page":"91","DOI":"10.1016\/j.ymeth.2019.02.009","article-title":"Deep-resp-forest: a deep forest model to predict anti-cancer drug response","volume":"166","author":"Ran","year":"2019","journal-title":"Methods"},{"issue":"1","key":"2022011920511078000_ref20","first-page":"1","article-title":"Functional random forest with applications in dose-response predictions","volume":"9","author":"Rahman","year":"2019","journal-title":"Sci Rep"},{"key":"2022011920511078000_ref21","doi-asserted-by":"crossref","first-page":"1041","DOI":"10.3389\/fgene.2019.01041","article-title":"Paclitaxel response can be predicted with interpretable multi-variate classifiers exploiting DNA-methylation and miRNA data","volume":"10","author":"Bomane","year":"2019","journal-title":"Front Genet"},{"key":"2022011920511078000_ref22","doi-asserted-by":"crossref","first-page":"509","DOI":"10.3389\/fchem.2019.00509","article-title":"Predicting synergism of cancer drug combinations using NCI-ALMANAC data","volume":"7","author":"Sidorov","year":"2019","journal-title":"Front Chem"},{"issue":"3","key":"2022011920511078000_ref23","doi-asserted-by":"crossref","first-page":"996","DOI":"10.1093\/bib\/bbz022","article-title":"Meta-GDBP: a high-level stacked regression model to improve anticancer drug response prediction","volume":"21","author":"Ran","year":"2020","journal-title":"Brief Bioinform"},{"issue":"24","key":"2022011920511078000_ref24","doi-asserted-by":"crossref","first-page":"5191","DOI":"10.1093\/bioinformatics\/btz418","article-title":"deepDR: a network-based deep learning approach to in silico drug repositioning","volume":"35","author":"Zeng","year":"2019","journal-title":"Bioinformatics"},{"key":"2022011920511078000_ref25","article-title":"DeepDSC: a deep learning method to predict drug sensitivity of cancer cell lines","volume":"18","author":"Li","year":"2019","journal-title":"IEEE\/ACM Trans Comput Biol Bioinform"},{"issue":"1","key":"2022011920511078000_ref26","first-page":"1","article-title":"A novel heterogeneous network-based method for drug response prediction in cancer cell lines","volume":"8","author":"Zhang","year":"2018","journal-title":"Sci Rep"},{"issue":"1","key":"2022011920511078000_ref27","first-page":"1","article-title":"Cancer drug response profile scan (CDRscan): a deep learning model that predicts drug effectiveness from cancer genomic signature","volume":"8","author":"Chang","year":"2018","journal-title":"Sci Rep"},{"issue":"1","key":"2022011920511078000_ref28","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1186\/s12859-019-2910-6","article-title":"Improving prediction of phenotypic drug response on cancer cell lines using deep convolutional network","volume":"20","author":"Liu","year":"2019","journal-title":"BMC Bioinformatics"},{"key":"2022011920511078000_ref29","article-title":"PaccMann: prediction of anticancer compound sensitivity with multi-modal attention-based neural networks","author":"Oskooei","year":"2018"},{"issue":"1","key":"2022011920511078000_ref30","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1038\/s41467-020-18197-y","article-title":"Representation of features as images with neighborhood dependencies for compatibility with convolutional neural networks","volume":"11","author":"Bazgir","year":"2020","journal-title":"Nat Commun"},{"issue":"2","key":"2022011920511078000_ref31","first-page":"325","article-title":"A review on machine learning principles for multi-view biological data integration","volume":"19","author":"Li","year":"2018","journal-title":"Brief Bioinform"},{"issue":"1","key":"2022011920511078000_ref32","doi-asserted-by":"crossref","first-page":"31","DOI":"10.1007\/s12551-018-0446-z","article-title":"Machine learning and feature selection for drug response prediction in precision oncology applications","volume":"11","author":"Ali","year":"2019","journal-title":"Biophys Rev"},{"issue":"2","key":"2022011920511078000_ref33","doi-asserted-by":"crossref","first-page":"671","DOI":"10.1093\/bib\/bby027","article-title":"Comparison and evaluation of integrative methods for the analysis of multilevel omics data: a study based on simulated and experimental cancer data","volume":"20","author":"Pucher","year":"2019","journal-title":"Brief Bioinform"},{"issue":"1","key":"2022011920511078000_ref34","first-page":"1","article-title":"Machine learning approaches to drug response prediction: challenges and recent progress","volume":"4","author":"Adam","year":"2020","journal-title":"NPJ Precis Oncol"},{"issue":"1","key":"2022011920511078000_ref35","doi-asserted-by":"crossref","first-page":"232","DOI":"10.1093\/bib\/bbz164","article-title":"A survey and systematic assessment of computational methods for drug response prediction","volume":"22","author":"Chen","year":"2021","journal-title":"Brief Bioinform"},{"issue":"4","key":"2022011920511078000_ref36","doi-asserted-by":"crossref","first-page":"749","DOI":"10.1002\/cpt.1773","article-title":"Machine learning for cancer drug combination","volume":"107","author":"Wang","year":"2020","journal-title":"Clin Pharmacol Therap"},{"issue":"1","key":"2022011920511078000_ref37","doi-asserted-by":"crossref","first-page":"360","DOI":"10.1093\/bib\/bbz171","article-title":"Deep learning for drug response prediction in cancer","volume":"22","author":"Baptista","year":"2021","journal-title":"Brief Bioinform"},{"issue":"1","key":"2022011920511078000_ref38","doi-asserted-by":"crossref","first-page":"346","DOI":"10.1093\/bib\/bbz153","article-title":"Improving drug response prediction by integrating multiple data sources: matrix factorization, kernel and network-based approaches","volume":"22","author":"Paltun","year":"2021","journal-title":"Brief Bioinform"},{"issue":"7480","key":"2022011920511078000_ref39","doi-asserted-by":"crossref","first-page":"389","DOI":"10.1038\/nature12831","article-title":"Inconsistency in large pharmacogenomic studies","volume":"504","author":"Haibe-Kains","year":"2013","journal-title":"Nature"},{"issue":"7631","key":"2022011920511078000_ref40","doi-asserted-by":"crossref","first-page":"E5","DOI":"10.1038\/nature20171","article-title":"Consistency in drug response profiling","volume":"540","author":"Mpindi","year":"2016","journal-title":"Nature"},{"key":"2022011920511078000_ref41","doi-asserted-by":"crossref","DOI":"10.12688\/f1000research.9611.1","article-title":"Revisiting inconsistency in large pharmacogenomic studies","volume":"5","author":"Safikhani","year":"2016","journal-title":"F1000Res"},{"issue":"1","key":"2022011920511078000_ref42","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1038\/s41598-017-14770-6","article-title":"New insight for pharmacogenomics studies from the transcriptional analysis of two large-scale cancer cell line panels","volume":"7","author":"Sadacca","year":"2017","journal-title":"Sci Rep"},{"issue":"7603","key":"2022011920511078000_ref43","doi-asserted-by":"crossref","first-page":"333","DOI":"10.1038\/nature17987","article-title":"Reproducible pharmacogenomic profiling of cancer cell line panels","volume":"533","author":"Haverty","year":"2016","journal-title":"Nature"},{"issue":"D1","key":"2022011920511078000_ref44","doi-asserted-by":"crossref","first-page":"D994","DOI":"10.1093\/nar\/gkx911","article-title":"Pharmacodb: an integrative database for mining in vitro anticancer drug screening studies","volume":"46","author":"Smirnov","year":"2017","journal-title":"Nucleic Acids Res"},{"issue":"5","key":"2022011920511078000_ref45","doi-asserted-by":"crossref","first-page":"1734","DOI":"10.1093\/bib\/bby046","article-title":"Evaluating the consistency of large-scale pharmacogenomic studies","volume":"20","author":"Rahman","year":"2019","journal-title":"Brief Bioinform"},{"issue":"1","key":"2022011920511078000_ref46","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1038\/s42003-020-0765-z","article-title":"A normalized drug response metric improves accuracy and consistency of anticancer drug sensitivity quantification in cell-based screening","volume":"3","author":"Gupta","year":"2020","journal-title":"Commun Biol"},{"issue":"17","key":"2022011920511078000_ref47","first-page":"51","article-title":"Application of transfer learning for cancer drug sensitivity prediction","volume":"19","author":"Dhruba","year":"2018","journal-title":"BMC Bioinformatics"},{"issue":"1","key":"2022011920511078000_ref48","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1038\/s41598-020-74921-0","article-title":"Ensemble transfer learning for the prediction of anti-cancer drug response","volume":"10","author":"Zhu","year":"2020","journal-title":"Sci Rep"},{"key":"2022011920511078000_ref49","article-title":"A systematic approach to featurization for cancer drug sensitivity predictions with deep learning","author":"Clyde","year":"2020"},{"key":"2022011920511078000_ref50","doi-asserted-by":"crossref","first-page":"5193","DOI":"10.1038\/srep05193","article-title":"Quantitative scoring of differential drug sensitivity for individually optimized anticancer therapies","volume":"4","author":"Yadav","year":"2014","journal-title":"Sci Rep"},{"issue":"1","key":"2022011920511078000_ref51","doi-asserted-by":"crossref","first-page":"5","DOI":"10.1023\/A:1010933404324","article-title":"Random forests","volume":"45","author":"Breiman","year":"2001","journal-title":"Mach Learn"},{"key":"2022011920511078000_ref52","first-page":"3146","article-title":"LightGBM: a highly efficient gradient boosting decision tree","volume":"30","author":"Ke","year":"2017","journal-title":"Adv Neural Inf Process Syst"},{"key":"2022011920511078000_ref53","article-title":"Learning curves for drug response prediction in cancer cell lines","volume-title":"BMC bioinformatics","author":"Partin","year":"2020"},{"issue":"3","key":"2022011920511078000_ref54","doi-asserted-by":"crossref","first-page":"985","DOI":"10.1093\/bib\/bbx153","article-title":"Effect of normalization methods on the performance of supervised learning algorithms applied to HTSeq-FPKM-UQ data sets: 7SK RNA expression as a predictor of survival in patients with colon adenocarcinoma","volume":"20","author":"Shahriyari","year":"2019","journal-title":"Brief Bioinform"},{"issue":"D1","key":"2022011920511078000_ref55","doi-asserted-by":"crossref","first-page":"D558","DOI":"10.1093\/nar\/gkx1063","article-title":"Data portal for the library of integrated network-based cellular signatures (LINCS) program: integrated access to diverse large-scale cellular perturbation response data","volume":"46","author":"Koleti","year":"2018","journal-title":"Nucleic Acids Res"},{"key":"2022011920511078000_ref56","volume-title":"Dragon (software for molecular descriptor calculation)","author":"Kode srl","year":"2017"},{"issue":"D1","key":"2022011920511078000_ref57","doi-asserted-by":"crossref","first-page":"D1102","DOI":"10.1093\/nar\/gky1033","article-title":"PubChem 2019 update: improved access to chemical data","volume":"47","author":"Kim","year":"2019","journal-title":"Nucleic Acids Res"},{"issue":"1","key":"2022011920511078000_ref58","first-page":"1","article-title":"Open babel: an open chemical toolbox","volume":"3","author":"O\u2019Boyle","year":"2011","journal-title":"J Chem"},{"key":"2022011920511078000_ref59","first-page":"2825","article-title":"Scikit-learn: machine learning in Python","volume":"12","author":"Pedregosa","year":"2011","journal-title":"J Mach Learn Res"},{"key":"2022011920511078000_ref60","first-page":"770","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"He","year":"2016"},{"issue":"2","key":"2022011920511078000_ref61","doi-asserted-by":"crossref","first-page":"90","DOI":"10.1038\/nchem.1243","article-title":"Quantifying the chemical beauty of drugs","volume":"4","author":"Bickerton","year":"2012","journal-title":"Nat Chem"},{"issue":"1A","key":"2022011920511078000_ref62","first-page":"A68","article-title":"The Cancer Genome Atlas (TCGA): an immeasurable source of knowledge","volume":"19","author":"Tomczak","year":"2015","journal-title":"Contemp Oncol"},{"issue":"4","key":"2022011920511078000_ref63","doi-asserted-by":"crossref","first-page":"453","DOI":"10.1182\/blood-2017-03-735654","article-title":"The NCI genomic data commons as an engine for precision medicine","volume":"130","author":"Jensen","year":"2017","journal-title":"Blood"},{"issue":"D1","key":"2022011920511078000_ref64","doi-asserted-by":"crossref","first-page":"D1100","DOI":"10.1093\/nar\/gkr777","article-title":"ChEMBL: a large-scale bioactivity database for drug discovery","volume":"40","author":"Gaulton","year":"2012","journal-title":"Nucleic Acids Res"},{"issue":"D1","key":"2022011920511078000_ref65","doi-asserted-by":"crossref","first-page":"D1045","DOI":"10.1093\/nar\/gkv1072","article-title":"BindingDB in 2015: a public database for medicinal chemistry, computational chemistry and systems pharmacology","volume":"44","author":"Gilson","year":"2016","journal-title":"Nucleic Acids Res"},{"issue":"8","key":"2022011920511078000_ref66","doi-asserted-by":"crossref","first-page":"1551","DOI":"10.1038\/nprot.2013.092","article-title":"Large-scale gene function analysis with the panther classification system","volume":"8","author":"Mi","year":"2013","journal-title":"Nat Protoc"},{"key":"2022011920511078000_ref67","first-page":"2672","article-title":"Generative adversarial nets","volume":"27","author":"Goodfellow","year":"2014","journal-title":"Adv Neural Inf Process Syst"},{"issue":"1","key":"2022011920511078000_ref68","doi-asserted-by":"crossref","first-page":"3","DOI":"10.1002\/(SICI)1098-1128(199601)16:1<3::AID-MED1>3.0.CO;2-6","article-title":"The art and practice of structure-based drug design: a molecular modeling perspective","volume":"16","author":"Bohacek","year":"1996","journal-title":"Med Res Rev"}],"container-title":["Briefings in Bioinformatics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/academic.oup.com\/bib\/article-pdf\/23\/1\/bbab356\/42229862\/bbab356.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/academic.oup.com\/bib\/article-pdf\/23\/1\/bbab356\/42229862\/bbab356.pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2023,11,8]],"date-time":"2023-11-08T14:24:17Z","timestamp":1699453457000},"score":1,"resource":{"primary":{"URL":"https:\/\/academic.oup.com\/bib\/article\/doi\/10.1093\/bib\/bbab356\/6370300"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2021,9,14]]},"references-count":68,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2022,1,17]]}},"URL":"https:\/\/doi.org\/10.1093\/bib\/bbab356","relation":{},"ISSN":["1467-5463","1477-4054"],"issn-type":[{"value":"1467-5463","type":"print"},{"value":"1477-4054","type":"electronic"}],"subject":[],"published-other":{"date-parts":[[2022,1]]},"published":{"date-parts":[[2021,9,14]]},"article-number":"bbab356"}}