{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,9]],"date-time":"2026-04-09T22:54:59Z","timestamp":1775775299028,"version":"3.50.1"},"reference-count":51,"publisher":"Oxford University Press (OUP)","issue":"3","license":[{"start":{"date-parts":[[2020,6,2]],"date-time":"2020-06-02T00:00:00Z","timestamp":1591056000000},"content-version":"vor","delay-in-days":0,"URL":"http:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"German Federal Ministry of Education and Research","award":["FKZ 03V0396"],"award-info":[{"award-number":["FKZ 03V0396"]}]},{"name":"European Union\u2019s Horizon 2020 research and innovation program","award":["633589"],"award-info":[{"award-number":["633589"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2021,5,20]]},"abstract":"<jats:title>Abstract<\/jats:title>\n               <jats:sec>\n                  <jats:title>Motivation<\/jats:title>\n                  <jats:p>The difficulty to find new drugs and bring them to the market has led to an increased interest to find new applications for known compounds. Biological samples from many disease contexts have been extensively profiled by transcriptomics, and, intuitively, this motivates to search for compounds with a reversing effect on the expression of characteristic disease genes. However, disease effects may be cell line-specific and also depend on other factors, such as genetics and environment. Transcription profile changes between healthy and diseased cells relate in complex ways to profile changes gathered from cell lines upon stimulation with a drug. Despite these differences, we expect that there will be some similarity in the gene regulatory networks at play in both situations. The challenge is to match transcriptomes for both diseases and drugs alike, even though the exact molecular pathology\/pharmacogenomics may not be known.<\/jats:p>\n               <\/jats:sec>\n               <jats:sec>\n                  <jats:title>Results<\/jats:title>\n                  <jats:p>We substitute the challenge to match a drug effect to a disease effect with the challenge to match a drug effect to the effect of the same drug at another concentration or in another cell line. This is welldefined, reproducible in vitro and in silico and extendable with external data. Based on the Connectivity Map (CMap) dataset, we combined 26 different similarity scores with six different heuristics to reduce the number of genes in the model. Such gene filters may also utilize external knowledge e.g. from biological networks. We found that no similarity score always outperforms all others for all drugs, but the Pearson correlation finds the same drug with the highest reliability. Results are improved by filtering for highly expressed genes and to a lesser degree for genes with large fold changes. Also a network-based reduction of contributing transcripts was beneficial, here implemented by the FocusHeuristics. We found no drop in prediction accuracy when reducing the whole transcriptome to the set of 1000 landmark genes of the CMap\u2019s successor project Library of Integrated Network-based Cellular Signatures. All source code to re-analyze and extend the CMap data, the source code of heuristics, filters and their evaluation are available to propel the development of new methods for drug repurposing.<\/jats:p>\n               <\/jats:sec>\n               <jats:sec>\n                  <jats:title>Availability<\/jats:title>\n                  <jats:p>https:\/\/bitbucket.org\/ibima\/moldrugeffectsdb<\/jats:p>\n               <\/jats:sec>\n               <jats:sec>\n                  <jats:title>Contact<\/jats:title>\n                  <jats:p>steffen.moeller@uni-rostock.de<\/jats:p>\n               <\/jats:sec>\n               <jats:sec>\n                  <jats:title>Supplementary information<\/jats:title>\n                  <jats:p>Supplementary data are available at Briefings in Bioinformatics online.<\/jats:p>\n               <\/jats:sec>","DOI":"10.1093\/bib\/bbaa072","type":"journal-article","created":{"date-parts":[[2020,4,20]],"date-time":"2020-04-20T11:09:27Z","timestamp":1587380967000},"source":"Crossref","is-referenced-by-count":22,"title":["Scoring functions for drug-effect similarity"],"prefix":"10.1093","volume":"22","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-8565-7962","authenticated-orcid":false,"given":"Stephan","family":"Struckmann","sequence":"first","affiliation":[{"name":"IBIMA, Rostock University Medical Center, Rostock, 18041, Germany"},{"name":"SHIP-KEF, Institute for Community Medicine, University Medicine of Greifswald, Walther-Rathenau-Stra\u03b2e 48, 17475 Greifswald, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-1096-4316","authenticated-orcid":false,"given":"Mathias","family":"Ernst","sequence":"additional","affiliation":[{"name":"IBIMA, Rostock University Medical Center, Rostock, 18041, Germany"},{"name":"Friedrich-Alexander-University Erlangen-Nuremberg, 91058 Erlangen, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-6218-7275","authenticated-orcid":false,"given":"Sarah","family":"Fischer","sequence":"additional","affiliation":[{"name":"IBIMA, Rostock University Medical Center, Rostock, 18041, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-1240-8076","authenticated-orcid":false,"given":"Nancy","family":"Mah","sequence":"additional","affiliation":[{"name":"BCRT - Berlin Institute of Health Center for Regenerative Therapies, Charit\u00e9 - University Medicine Berlin, 13353, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-4994-9829","authenticated-orcid":false,"given":"Georg","family":"Fuellen","sequence":"additional","affiliation":[{"name":"IBIMA, Rostock University Medical Center, Rostock, 18041, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-7187-4683","authenticated-orcid":false,"given":"Steffen","family":"M\u00f6ller","sequence":"additional","affiliation":[{"name":"IBIMA, Rostock University Medical Center, Rostock, 18041, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"286","published-online":{"date-parts":[[2020,6,2]]},"reference":[{"issue":"12","key":"2021052109430247200_ref1","doi-asserted-by":"crossref","first-page":"R139","DOI":"10.1186\/gb-2009-10-12-r139","article-title":"Mining for coexpression across hundreds of datasets using novel rank aggregation and visualization methods","volume":"10","author":"Adler","year":"2009","journal-title":"Genome Biol"},{"key":"2021052109430247200_ref2","first-page":"bts707","article-title":"Minerva and minepy: a C engine for the mine suite and its R, Python and matlab wrappers","volume-title":"Bioinformatics","author":"Albanese","year":"2012"},{"issue":"D1","key":"2021052109430247200_ref3","doi-asserted-by":"crossref","first-page":"D711","DOI":"10.1093\/nar\/gky964","article-title":"ArrayExpress update\u2014from bulk to single-cell expression data","volume":"47","author":"Athar","year":"2018","journal-title":"Nucleic Acids Res"},{"issue":"22","key":"2021052109430247200_ref4","doi-asserted-by":"crossref","first-page":"2641","DOI":"10.4155\/fmc-2018-0076","article-title":"A review of ligand-based virtual screening web tools and screening algorithms in large molecular databases in the age of big data","volume":"10","author":"Banegas-Luna","year":"2018","journal-title":"Future Med Chem"},{"key":"2021052109430247200_ref5","doi-asserted-by":"crossref","first-page":"D991","DOI":"10.1093\/nar\/gks1193","article-title":"NCBI GEO: archive for functional genomics data sets\u2014update","volume":"41","author":"Barrett","year":"2013","journal-title":"Nucleic Acids Res"},{"key":"2021052109430247200_ref6","doi-asserted-by":"crossref","first-page":"2337","DOI":"10.2174\/15680266113136660164","article-title":"Drug repositioning for treatment of movement disorders: from serendipity to rational discovery strategies","volume":"13","author":"Bolg\u00e1r","year":"2013","journal-title":"Curr Top Med Chem"},{"key":"2021052109430247200_ref7","doi-asserted-by":"crossref","first-page":"170029","DOI":"10.1038\/sdata.2017.29","article-title":"A standard database for drug repositioning","volume":"4","author":"Brown","year":"2017","journal-title":"Sci Data"},{"key":"2021052109430247200_ref8","article-title":"org.Hs.eg.db: Genome Wide Annotation for Human","author":"Carlson","year":"2016"},{"issue":"16","key":"2021052109430247200_ref9","doi-asserted-by":"crossref","first-page":"2818","DOI":"10.1093\/bioinformatics\/btz006","article-title":"Breaking the paradigm: Dr insight empowers signature-free, enhanced drug repurposing","volume":"35","author":"Chan","year":"2019","journal-title":"Bioinformatics"},{"issue":"D1","key":"2021052109430247200_ref10","doi-asserted-by":"crossref","first-page":"D948","DOI":"10.1093\/nar\/gky868","article-title":"The Comparative Toxicogenomics Database: update 2019","volume":"47","author":"Davis","year":"2019","journal-title":"Nucleic Acids Res"},{"issue":"5","key":"2021052109430247200_ref11","doi-asserted-by":"crossref","DOI":"10.1111\/acel.12819","article-title":"Gene expression-based drug repurposing to target aging","volume":"17","author":"D\u00f6nerta\u015f","year":"2018","journal-title":"Aging Cell"},{"key":"2021052109430247200_ref12","doi-asserted-by":"crossref","first-page":"W449","DOI":"10.1093\/nar\/gku476","article-title":"Lincs canvas browser: interactive web app to query, browse and interrogate lincs l1000 gene expression signatures","volume":"42","author":"Duan","year":"2014","journal-title":"Nucleic Acids Res"},{"issue":"16","key":"2021052109430247200_ref13","doi-asserted-by":"crossref","first-page":"3439","DOI":"10.1093\/bioinformatics\/bti525","article-title":"BioMart and Bioconductor: a powerful link between biological databases and microarray data analysis","volume":"21","author":"Durinck","year":"2005","journal-title":"Bioinformatics"},{"issue":"8","key":"2021052109430247200_ref14","doi-asserted-by":"crossref","first-page":"1184","DOI":"10.1038\/nprot.2009.97","article-title":"Mapping identifiers for the integration of genomic datasets with the R\/Bioconductor package biomaRt","volume":"4","author":"Durinck","year":"2009","journal-title":"Nat Protoc"},{"issue":"5","key":"2021052109430247200_ref15","doi-asserted-by":"crossref","first-page":"1027","DOI":"10.1002\/asi.21009","article-title":"The relation between Pearson\u2019s correlation coefficient r and Salton\u2019s cosine measure","volume":"60","author":"Egghe","year":"2009","journal-title":"J Am Soc Inf Sci Technol"},{"key":"2021052109430247200_ref16","doi-asserted-by":"crossref","first-page":"42638","DOI":"10.1038\/srep42638","article-title":"FocusHeuristics\u2014expression-data-driven network optimization and disease gene prediction","volume":"7","author":"Ernst","year":"2017","journal-title":"Sci Rep"},{"key":"2021052109430247200_ref17","article-title":"String v9.1: protein\u2013protein interaction networks, with increased coverage and integration.","volume-title":"Nucleic Acids Res.","author":"Franceschini","year":"2013"},{"issue":"3","key":"2021052109430247200_ref18","doi-asserted-by":"crossref","first-page":"219","DOI":"10.1016\/j.jbiotec.2005.03.022","article-title":"Development of a large-scale chemogenomics database to improve drug candidate selection and to understand mechanisms of chemical toxicity and action","volume":"119","author":"Ganter","year":"2005","journal-title":"J Biotechnol"},{"issue":"D1","key":"2021052109430247200_ref19","doi-asserted-by":"crossref","first-page":"D945","DOI":"10.1093\/nar\/gkw1074","article-title":"The ChEMBL database in 2017","volume":"45","author":"Gaulton","year":"2017","journal-title":"Nucleic Acids Res"},{"key":"2021052109430247200_ref20","article-title":"CDEK: Clinical Drug Experience Knowledgebase","volume-title":"Database","author":"Griesenauer","year":"2019"},{"key":"2021052109430247200_ref21","article-title":"biwt: Functions to Compute the Biweight Mean Vector and Covariance & Correlation Matrices","author":"Hardin","year":"2012"},{"key":"2021052109430247200_ref22","doi-asserted-by":"crossref","first-page":"D921","DOI":"10.1093\/nar\/gku955","article-title":"Open TG-GATEs: a large-scale toxicogenomics database","volume":"43","author":"Igarashi","year":"2015","journal-title":"Nucleic Acids Res"},{"key":"2021052109430247200_ref23","doi-asserted-by":"crossref","DOI":"10.1093\/nar\/gkz957","article-title":"TSEA-DB: a trait-tissue association map for human complex traits and diseases","author":"Jia","year":"2019","journal-title":"Nucleic Acids Res."},{"issue":"1","key":"2021052109430247200_ref24","doi-asserted-by":"crossref","first-page":"414","DOI":"10.1186\/s12864-016-2737-8","article-title":"Cogena, a novel tool for co-expressed gene-set enrichment analysis, applied to drug repositioning and drug mode of action discovery","volume":"17","author":"Jia","year":"2016","journal-title":"BMC Genom"},{"issue":"1","key":"2021052109430247200_ref25","doi-asserted-by":"crossref","first-page":"69","DOI":"10.1146\/annurev-biodatasci-072018-021211","article-title":"Connectivity mapping: methods and applications","volume":"2","author":"Keenan","year":"2019","journal-title":"Annu Rev Biomed Data Sci"},{"issue":"D1","key":"2021052109430247200_ref26","doi-asserted-by":"crossref","first-page":"D1202","DOI":"10.1093\/nar\/gkv951","article-title":"PubChem substance and compound databases","volume":"44","author":"Kim","year":"2016","journal-title":"Nucleic Acids Res"},{"key":"2021052109430247200_ref27","first-page":"bar030","article-title":"Ensembl BioMarts: a hub for data retrieval across taxonomic space","volume-title":"Database (Oxford)","author":"Kinsella","year":"2011"},{"issue":"W1","key":"2021052109430247200_ref28","doi-asserted-by":"crossref","first-page":"W183","DOI":"10.1093\/nar\/gkz347","article-title":"modEnrichr: a suite of gene set enrichment analysis tools for model organisms","volume":"47","author":"Kuleshov","year":"2019","journal-title":"Nucleic Acids Res"},{"issue":"5795","key":"2021052109430247200_ref29","doi-asserted-by":"crossref","first-page":"1929","DOI":"10.1126\/science.1132939","article-title":"The connectivity map: using gene-expression signatures to connect small molecules, genes, and disease","volume":"313","author":"Lamb","year":"2006","journal-title":"Science"},{"key":"2021052109430247200_ref30","doi-asserted-by":"crossref","first-page":"559","DOI":"10.1186\/1471-2105-9-559","article-title":"WGCNA: an R package for weighted correlation network analysis","volume":"9","author":"Langfelder","year":"2008","journal-title":"BMC Bioinform"},{"key":"2021052109430247200_ref31","article-title":"sva: Surrogate Variable Analysis","author":"Leek","year":"2016"},{"issue":"1","key":"2021052109430247200_ref32","doi-asserted-by":"crossref","first-page":"2","DOI":"10.1093\/bib\/bbv020","article-title":"A survey of current trends in computational drug repositioning","volume":"17","author":"Li","year":"2015","journal-title":"Brief Bioinform"},{"key":"2021052109430247200_ref33","article-title":"rsgcc: Gini Methodology-Based Correlation and Clustering Analysis of Microarray and RNA-Seq Gene Expression Data","author":"Ma","year":"2013"},{"issue":"6","key":"2021052109430247200_ref34","doi-asserted-by":"crossref","DOI":"10.1371\/journal.pcbi.1005487","article-title":"Comorbidities in the diseasome are more apparent than real: what Bayesian filtering reveals about the comorbidities of depression","volume":"13","author":"Marx","year":"2017","journal-title":"PLoS Comput Biol"},{"key":"2021052109430247200_ref35","doi-asserted-by":"crossref","first-page":"250","DOI":"10.1007\/978-3-030-16272-6_9","article-title":"Cloud-based high throughput virtual screening in novel drug discovery","volume-title":"High-Performance Modelling and Simulation for Big Data Applications","author":"Ol\u011fa\u00e7","year":"2019"},{"key":"2021052109430247200_ref36","first-page":"e1006651","article-title":"Predicting protein targets for drug-like compounds using transcriptomics.","volume-title":"PLoS Comput Biol","author":"Pabon","year":"2018"},{"key":"2021052109430247200_ref37","volume-title":"R: A Language and Environment for Statistical Computing","author":"Core Team","year":"2016"},{"key":"2021052109430247200_ref38","doi-asserted-by":"crossref","DOI":"10.1093\/nar\/gkz369","article-title":"g:profiler: a web server for functional enrichment analysis and conversions of gene lists (2019 update)","author":"Raudvere","year":"2019","journal-title":"Nucleic Acids Res"},{"issue":"5","key":"2021052109430247200_ref39","doi-asserted-by":"crossref","first-page":"1878","DOI":"10.1093\/bib\/bby061","article-title":"Recent applications of deep learning and machine intelligence on in silico drug discovery: methods, tools and databases","volume":"20","author":"Rifaioglu","year":"2018","journal-title":"Brief Bioinform"},{"key":"2021052109430247200_ref40","article-title":"parmigene: Parallel Mutual Information estimation for Gene Network reconstruction.","author":"Sales","year":"2012"},{"issue":"W1","key":"2021052109430247200_ref41","doi-asserted-by":"crossref","first-page":"W193","DOI":"10.1093\/nar\/gkv445","article-title":"NFFinder: an online bioinformatics tool for searching similar transcriptomics experiments in the context of drug repositioning","volume":"43","author":"Setoain","year":"2015","journal-title":"Nucleic Acids Res"},{"issue":"1","key":"2021052109430247200_ref42","doi-asserted-by":"crossref","first-page":"225","DOI":"10.1186\/s13023-019-1193-3","article-title":"The use or generation of biomedical data and existing medicines to discover and establish new treatments for patients with rare diseases\u2014recommendations of the IRDiRC Data Mining and Repurposing Task Force","volume":"14","author":"Southall","year":"2019","journal-title":"Orphanet J Rare Dis"},{"issue":"6","key":"2021052109430247200_ref43","doi-asserted-by":"crossref","first-page":"1437","DOI":"10.1016\/j.cell.2017.10.049","article-title":"A next generation connectivity map: L1000 platform and the first 1,000,000 profiles","volume":"171","author":"Subramanian","year":"2017","journal-title":"Cell"},{"key":"2021052109430247200_ref44","doi-asserted-by":"crossref","first-page":"D447","DOI":"10.1093\/nar\/gku1003","article-title":"STref37ING v10: protein\u2013protein interaction networks, integrated over the tree of life","volume":"43","author":"Szklarczyk","year":"2015","journal-title":"Nucleic Acids Res."},{"key":"2021052109430247200_ref45","doi-asserted-by":"publisher","DOI":"10.1101\/843029","article-title":"DeepSide: a deep learning framework for drug side effect prediction","author":"Uner","year":"2019"},{"issue":"1","key":"2021052109430247200_ref46","doi-asserted-by":"crossref","first-page":"10","DOI":"10.1093\/nar\/28.1.10","article-title":"Database resources of the national center for biotechnology information","volume":"28","author":"Wheeler","year":"2000","journal-title":"Nucleic Acids Res"},{"key":"2021052109430247200_ref47","doi-asserted-by":"crossref","DOI":"10.1007\/978-3-319-24277-4","volume-title":"ggplot2: Elegant Graphics for Data Analysis","author":"Wickham","year":"2016"},{"key":"2021052109430247200_ref48","doi-asserted-by":"crossref","first-page":"160018","DOI":"10.1038\/sdata.2016.18","article-title":"The FAIR Guiding Principles for scientific data management and stewardship","volume":"3","author":"Wilkinson","year":"2016","journal-title":"Sci Data"},{"issue":"7","key":"2021052109430247200_ref49","doi-asserted-by":"crossref","first-page":"667","DOI":"10.1186\/s12864-018-5031-0","article-title":"Deep learning-based transcriptome data classification for drug\u2013target interaction prediction","volume":"19","author":"Xie","year":"2018","journal-title":"BMC Genom"},{"issue":"24","key":"2021052109430247200_ref50","doi-asserted-by":"crossref","first-page":"5191","DOI":"10.1093\/bioinformatics\/btz418","article-title":"deepDR: a network-based deep learning approach to in silico drug repositioning","volume":"35","author":"Zeng","year":"2019","journal-title":"Bioinformatics"},{"key":"2021052109430247200_ref51","doi-asserted-by":"crossref","first-page":"1171","DOI":"10.1038\/s41588-018-0160-6","article-title":"Deep learning sequence-based ab initio prediction of variant effects on expression and disease risk","volume":"50","author":"Zhou","year":"2018","journal-title":"Nat Genet"}],"container-title":["Briefings in Bioinformatics"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/academic.oup.com\/bib\/article-pdf\/22\/3\/bbaa072\/37965803\/bbaa072.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"http:\/\/academic.oup.com\/bib\/article-pdf\/22\/3\/bbaa072\/37965803\/bbaa072.pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2021,5,21]],"date-time":"2021-05-21T09:44:00Z","timestamp":1621590240000},"score":1,"resource":{"primary":{"URL":"https:\/\/academic.oup.com\/bib\/article\/doi\/10.1093\/bib\/bbaa072\/5850229"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,6,2]]},"references-count":51,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2021,5,20]]}},"URL":"https:\/\/doi.org\/10.1093\/bib\/bbaa072","relation":{},"ISSN":["1467-5463","1477-4054"],"issn-type":[{"value":"1467-5463","type":"print"},{"value":"1477-4054","type":"electronic"}],"subject":[],"published-other":{"date-parts":[[2021,5]]},"published":{"date-parts":[[2020,6,2]]},"article-number":"bbaa072"}}