{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,15]],"date-time":"2026-07-15T07:15:00Z","timestamp":1784099700498,"version":"3.55.0"},"reference-count":57,"publisher":"Springer Science and Business Media LLC","issue":"4","license":[{"start":{"date-parts":[[2024,2,13]],"date-time":"2024-02-13T00:00:00Z","timestamp":1707782400000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2024,2,13]],"date-time":"2024-02-13T00:00:00Z","timestamp":1707782400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"DOI":"10.13039\/501100006764","name":"Technische Universit\u00e4t Berlin","doi-asserted-by":"crossref","id":[{"id":"10.13039\/501100006764","id-type":"DOI","asserted-by":"crossref"}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["The VLDB Journal"],"published-print":{"date-parts":[[2024,7]]},"abstract":"<jats:title>Abstract<\/jats:title><jats:p>When designing data science (DS) pipelines, end-users can get overwhelmed by the large and growing set of available data preprocessing and modeling techniques. Intelligent discovery assistants (IDAs) and automated machine learning (AutoML) solutions aim to facilitate end-users by (semi-)automating the process. However, they are expensive to compute and yield limited applicability for a wide range of real-world use cases and application domains. This is due to (a) their need to execute thousands of pipelines to get the optimal one, (b) their limited support of DS tasks, e.g., supervised classification or regression only, and a small, static set of available data preprocessing and ML algorithms; and (c) their restriction to quantifiable evaluation processes and metrics, e.g., tenfold cross-validation using the ROC AUC score for classification. To overcome these limitations, we propose a human-in-the-loop approach for the<jats:italic>assisted<\/jats:italic><jats:italic>design<\/jats:italic><jats:italic>of<\/jats:italic><jats:italic>data<\/jats:italic><jats:italic>science<\/jats:italic><jats:italic>pipelines<\/jats:italic>using previously executed pipelines. Based on a user query, i.e.,\u00a0data and a DS task, our framework outputs a ranked list of pipeline candidates from which the user can choose to execute or modify in real time. To recommend pipelines, it first identifies relevant datasets and pipelines utilizing efficient similarity search. It then ranks the candidate pipelines using multi-objective sorting and takes user interactions into account to improve suggestions over time. In our experimental evaluation, the proposed framework significantly outperforms the state-of-the-art IDA tool and achieves similar predictive performance with state-of-the-art long-running AutoML solutions while being real-time, generic to any evaluation processes and DS tasks, and extensible to new operators.<\/jats:p>","DOI":"10.1007\/s00778-024-00835-2","type":"journal-article","created":{"date-parts":[[2024,2,13]],"date-time":"2024-02-13T09:03:27Z","timestamp":1707815007000},"page":"1129-1153","update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":7,"title":["Assisted design of data science pipelines"],"prefix":"10.1007","volume":"33","author":[{"ORCID":"https:\/\/orcid.org\/0000-0001-7131-745X","authenticated-orcid":false,"given":"Sergey","family":"Redyuk","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Zoi","family":"Kaoudi","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Sebastian","family":"Schelter","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Volker","family":"Markl","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2024,2,13]]},"reference":[{"issue":"1","key":"835_CR1","doi-asserted-by":"publisher","first-page":"39","DOI":"10.3233\/AIC-1994-7104","volume":"7","author":"A Aamodt","year":"1994","unstructured":"Aamodt, A., Plaza, E.: Case-based reasoning: foundational issues, methodological variations, and system approaches. AI Commun. 7(1), 39\u201359 (1994)","journal-title":"AI Commun."},{"key":"835_CR2","doi-asserted-by":"publisher","unstructured":"Abu-Aisheh, Z., Raveaux, R., Ramel, J., Martineau, P.: An exact graph edit distance algorithm for solving pattern recognition problems. In: ICPRAM\u201915. Lisbon, Portugal (2015). https:\/\/doi.org\/10.5220\/0005209202710278, https:\/\/hal.archives-ouvertes.fr\/hal-01168816","DOI":"10.5220\/0005209202710278"},{"key":"835_CR3","unstructured":"Amashukeli, S., Elshawi, R., Sakr, S.: ismartml: an interactive and user-guided framework for automated machine learning. In: Proceedings of the Workshop on Human-In-the-Loop Data Analytics, HILDA\u201920 (2020)"},{"issue":"6","key":"835_CR4","doi-asserted-by":"publisher","first-page":"592","DOI":"10.1038\/s41587-019-0140-0","volume":"37","author":"\u017d Avsec","year":"2019","unstructured":"Avsec, \u017d, et al.: The kipoi repository accelerates community exchange and reuse of predictive models for genomics. Nat. Biotechnol. 37(6), 592\u2013600 (2019)","journal-title":"Nat. Biotechnol."},{"issue":"9","key":"835_CR5","doi-asserted-by":"publisher","first-page":"509","DOI":"10.1145\/361002.361007","volume":"18","author":"JL Bentley","year":"1975","unstructured":"Bentley, J.L.: Multidimensional binary search trees used for associative searching. Commun. ACM 18(9), 509\u2013517 (1975). https:\/\/doi.org\/10.1145\/361002.361007","journal-title":"Commun. ACM"},{"key":"835_CR6","doi-asserted-by":"crossref","unstructured":"Bergstra, J., Yamins, D., Cox, D.: Hyperopt: a python library for optimizing the hyperparameters of machine learning algorithms. In: SciPy\u201913, vol.\u00a013, p.\u00a020. Citeseer (2013)","DOI":"10.25080\/Majora-8b375195-003"},{"key":"835_CR7","doi-asserted-by":"publisher","first-page":"101","DOI":"10.1016\/j.csi.2017.05.004","volume":"57","author":"B Bilalli","year":"2018","unstructured":"Bilalli, B., Abell\u00f3, A., Aluja-Banet, T., Wrembel, R.: Intelligent assistance for data pre-processing. Comput. Stand. Interfaces 57, 101\u2013109 (2018)","journal-title":"Comput. Stand. Interfaces"},{"key":"835_CR8","unstructured":"Bischl, B., et\u00a0al.: Openml benchmarking suites and the openml100. stat 1050, 11 (2017)"},{"key":"835_CR9","unstructured":"Borges, R., Stefanidis, K.: On measuring popularity bias in collaborative filtering data. In: EDBT\/ICDT Workshops (2020)"},{"key":"835_CR10","doi-asserted-by":"publisher","unstructured":"Brazdil, P., van Rijn, J., Soares, C., Vanschoren, J.: Automating workflow\/pipeline design, pp. 123\u2013140. Springer, Cham (2022). https:\/\/doi.org\/10.1007\/978-3-030-67024-5_7","DOI":"10.1007\/978-3-030-67024-5_7"},{"issue":"4","key":"835_CR11","doi-asserted-by":"publisher","first-page":"230","DOI":"10.1145\/362003.362025","volume":"16","author":"WA Burkhard","year":"1973","unstructured":"Burkhard, W.A., Keller, R.M.: Some approaches to best-match file searching. Commun. ACM 16(4), 230\u2013236 (1973). https:\/\/doi.org\/10.1145\/362003.362025","journal-title":"Commun. ACM"},{"key":"835_CR12","doi-asserted-by":"crossref","unstructured":"Buzdalov, M., Shalyto, A.: A provably asymptotically fast version of the generalized jensen algorithm for non-dominated sorting. In: PPSN\u201914, pp. 528\u2013537. Springer (2014)","DOI":"10.1007\/978-3-319-10762-2_52"},{"key":"835_CR13","doi-asserted-by":"publisher","unstructured":"Cambronero, J.P., Rinard, M.C.: Al: Autogenerating supervised learning programs. Proc. ACM Program. Lang. 3(OOPSLA) (2019). https:\/\/doi.org\/10.1145\/3360601","DOI":"10.1145\/3360601"},{"issue":"2","key":"835_CR14","first-page":"10","volume":"41","author":"C Chen","year":"2018","unstructured":"Chen, C., Golshan, B., Halevy, A.Y., Tan, W.C., Doan, A.: Biggorilla: an open-source ecosystem for data preparation and integration. IEEE Data Eng. Bull. 41(2), 10\u201322 (2018)","journal-title":"IEEE Data Eng. Bull."},{"key":"835_CR15","doi-asserted-by":"publisher","unstructured":"Cordella, L., Foggia, P., Sansone, C., Vento, M.: Performance evaluation of the vf graph matching algorithm. In: Proceedings 10th International Conference on Image Analysis and Processing, pp. 1172\u20131177 (1999). https:\/\/doi.org\/10.1109\/ICIAP.1999.797762","DOI":"10.1109\/ICIAP.1999.797762"},{"issue":"10","key":"835_CR16","doi-asserted-by":"publisher","first-page":"1367","DOI":"10.1109\/TPAMI.2004.75","volume":"26","author":"L Cordella","year":"2004","unstructured":"Cordella, L., Foggia, P., Sansone, C., Vento, M.: A (sub) graph isomorphism algorithm for matching large graphs. IEEE Trans. Pattern Anal. Mach. Intell. 26(10), 1367\u20131372 (2004)","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"835_CR17","doi-asserted-by":"crossref","unstructured":"Corradini, A., Heindel, T., Hermann, F., K\u00f6nig, B.: Sesqui-pushout rewriting. In: International Conference on Graph Transformation, pp. 30\u201345. Springer (2006)","DOI":"10.1007\/11841883_4"},{"key":"835_CR18","doi-asserted-by":"crossref","unstructured":"Craw, S., Sleeman, D., Graner, N., Rissakis, M., Sharma, S.: Consultant: providing advice for the machine learning toolbox. In: Proceedings of the Research and Development in Expert Systems IX, pp. 5\u201323 (1992)","DOI":"10.1017\/CBO9780511569944.002"},{"issue":"1","key":"835_CR19","doi-asserted-by":"publisher","first-page":"86","DOI":"10.1016\/S0022-0000(75)80051-1","volume":"11","author":"A Cremers","year":"1975","unstructured":"Cremers, A., Ginsburg, S.: Context-free grammar forms. J. Comput. Syst. Sci. 11(1), 86\u2013117 (1975)","journal-title":"J. Comput. Syst. Sci."},{"issue":"3","key":"835_CR20","first-page":"37","volume":"17","author":"U Fayyad","year":"1996","unstructured":"Fayyad, U., Piatetsky-Shapiro, G., Smyth, P.: From data mining to knowledge discovery in databases. AI Mag. 17(3), 37\u201337 (1996)","journal-title":"AI Mag."},{"key":"835_CR21","doi-asserted-by":"crossref","unstructured":"Feurer, M., et\u00a0al.: Auto-sklearn: efficient and robust automated machine learning. In: Automated Machine Learning, pp. 113\u2013134. Springer, Cham (2019)","DOI":"10.1007\/978-3-030-05318-5_6"},{"key":"835_CR22","unstructured":"Fusi, N., Sheth, R., Elibol, M.: Probabilistic matrix factorization for automated machine learning. In: Advances in Neural Information Processing Systems, vol. 31 (2018)"},{"key":"835_CR23","doi-asserted-by":"publisher","unstructured":"He, X., Zhao, K., Chu, X.: Automl: a survey of the state-of-the-art. Knowledge-Based Systems 212, 106,622 (2021). https:\/\/doi.org\/10.1016\/j.knosys.2020.106622, https:\/\/www.sciencedirect.com\/science\/article\/pii\/S0950705120307516","DOI":"10.1016\/j.knosys.2020.106622"},{"key":"835_CR24","volume-title":"Ansible: Up and Running: Automating Configuration Management and Deployment the Easy Way","author":"L Hochstein","year":"2017","unstructured":"Hochstein, L., Moser, R.: Ansible: Up and Running: Automating Configuration Management and Deployment the Easy Way. O\u2019Reilly Media Inc, New York (2017)"},{"issue":"5","key":"835_CR25","doi-asserted-by":"publisher","first-page":"503","DOI":"10.1109\/TEVC.2003.817234","volume":"7","author":"M Jensen","year":"2003","unstructured":"Jensen, M.: Reducing the run-time complexity of multiobjective eas: The nsga-ii and other algorithms. IEEE Trans. Evol. Comput. 7(5), 503\u2013515 (2003). https:\/\/doi.org\/10.1109\/TEVC.2003.817234","journal-title":"IEEE Trans. Evol. Comput."},{"issue":"1","key":"835_CR26","first-page":"826","volume":"18","author":"L Kotthoff","year":"2017","unstructured":"Kotthoff, L., Thornton, C., Hoos, H., Hutter, F., Leyton-Brown, K.: Auto-weka 2.0: automatic model selection and hyperparameter optimization in weka. J. Mach. Learn. Res. 18(1), 826\u2013830 (2017)","journal-title":"J. Mach. Learn. Res."},{"issue":"1","key":"835_CR27","doi-asserted-by":"publisher","first-page":"250","DOI":"10.1093\/bioinformatics\/btz470","volume":"36","author":"TT Le","year":"2020","unstructured":"Le, T.T., Fu, W., Moore, J.H.: Scaling tree-based automated machine learning to biomedical big data with a feature set selector. Bioinformatics 36(1), 250\u2013256 (2020)","journal-title":"Bioinformatics"},{"key":"835_CR28","unstructured":"Liaw, R., et\u00a0al.: Tune: a research platform for distributed model selection and training. arXiv:1807.05118 (2018)"},{"key":"835_CR29","doi-asserted-by":"publisher","unstructured":"Lika, B., Kolomvatsos, K., Hadjiefthymiades, S.: Facing the cold start problem in recommender systems. Expert Systems with Applications 41(4, Part 2), 2065\u20132073 (2014). https:\/\/doi.org\/10.1016\/j.eswa.2013.09.005, https:\/\/www.sciencedirect.com\/science\/article\/pii\/S0957417413007240","DOI":"10.1016\/j.eswa.2013.09.005"},{"key":"835_CR30","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-14267-3","volume-title":"Learning to Rank for Information Retrieval","author":"TY Liu","year":"2011","unstructured":"Liu, T.Y.: Learning to Rank for Information Retrieval. Springer, Berlin (2011)"},{"issue":"3","key":"835_CR31","doi-asserted-by":"publisher","first-page":"225","DOI":"10.1561\/1500000016","volume":"3","author":"TY Liu","year":"2009","unstructured":"Liu, T.Y., et al.: Learning to rank for information retrieval. Found. Trends Inf. Retriev. 3(3), 225\u2013331 (2009)","journal-title":"Found. Trends Inf. Retriev."},{"issue":"6","key":"835_CR32","doi-asserted-by":"publisher","DOI":"10.15252\/msb.20188746","volume":"15","author":"M Luecken","year":"2019","unstructured":"Luecken, M., Theis, F.: Current best practices in single-cell rna-seq analysis: a tutorial. Mol. Syst. Biol 15(6), e8746 (2019)","journal-title":"Mol. Syst. Biol"},{"key":"835_CR33","first-page":"30","volume":"87","author":"B McKay","year":"1981","unstructured":"McKay, B.: Practical graph isomorphism. Congr. Numerantium 87, 30\u201345 (1981)","journal-title":"Congr. Numerantium"},{"issue":"2","key":"835_CR34","doi-asserted-by":"publisher","first-page":"81","DOI":"10.1037\/h0043158","volume":"63","author":"GA Miller","year":"1956","unstructured":"Miller, G.A.: The magical number seven, plus or minus two: some limits on our capacity for processing information. Psychol. Rev. 63(2), 81 (1956)","journal-title":"Psychol. Rev."},{"key":"835_CR35","doi-asserted-by":"publisher","unstructured":"Miller, R.B.: Response time in man-computer conversational transactions. In: Proceedings of the December 9\u201311, 1968, Fall Joint Computer Conference, Part I, AFIPS \u201968 (Fall, part I), p. 267-277. ACM, New York (1968). https:\/\/doi.org\/10.1145\/1476589.1476628","DOI":"10.1145\/1476589.1476628"},{"key":"835_CR36","doi-asserted-by":"publisher","unstructured":"M\u00f6lder, F., Jablonski, K., Letcher, B., Hall, M., Tomkins-Tinch, C., Sochat, V., Forster, J., Lee, S., Twardziok, S., Kanitz, A., Wilm, A., Holtgrewe, M., Rahmann, S., Nahnsen, S., K\u00f6ster, J.: Sustainable data analysis with snakemake [version 2; peer review: 2 approved]. F1000Research 10(33) (2021). https:\/\/doi.org\/10.12688\/f1000research.29032.2","DOI":"10.12688\/f1000research.29032.2"},{"key":"835_CR37","doi-asserted-by":"publisher","unstructured":"Namaki, M.H., Floratou, A., Psallidas, F., Krishnan, S., Agrawal, A., Wu, Y., Zhu, Y., Weimer, M.: Vamsa: automated provenance tracking in data science scripts. In: Proceedings of the 26th ACM SIGKDD International Conference on Knowledge Discovery and Data Mining, KDD \u201920, pp. 1542\u20131551. Association for Computing Machinery, New York (2020). https:\/\/doi.org\/10.1145\/3394486.3403205","DOI":"10.1145\/3394486.3403205"},{"issue":"1","key":"835_CR38","first-page":"605","volume":"51","author":"P Nguyen","year":"2014","unstructured":"Nguyen, P., Hilario, M., Kalousis, A.: Using meta-mining to support data mining workflow planning and optimization. J. Artif. Int. Res. 51(1), 605\u2013644 (2014)","journal-title":"J. Artif. Int. Res."},{"key":"835_CR39","unstructured":"Olson, R., Moore, J.: Tpot: a tree-based pipeline optimization tool for automating machine learning. In: ICML\u201916 AutoML Workshop, pp. 66\u201374. JMLR (2016)"},{"key":"835_CR40","doi-asserted-by":"crossref","unstructured":"Patterson, E., Baldini, I., Mojsilovic, A., Varshney, K.R.: Semantic representation of data science programs. In: IJCAI, pp. 5847\u20135849 (2018)","DOI":"10.24963\/ijcai.2018\/858"},{"key":"835_CR41","doi-asserted-by":"publisher","unstructured":"Rahman, S., Rochan, M.: A fast farthest neighbor search algorithm for very high dimensional data. In: 19th International Conference on Computer and Information Technology (ICCIT), pp. 351\u2013356 (2016). https:\/\/doi.org\/10.1109\/ICCITECHN.2016.7860222","DOI":"10.1109\/ICCITECHN.2016.7860222"},{"issue":"12","key":"835_CR42","doi-asserted-by":"publisher","first-page":"3714","DOI":"10.14778\/3554821.3554882","volume":"15","author":"S Redyuk","year":"2022","unstructured":"Redyuk, S., Kaoudi, Z., Schelter, S., Markl, V.: DORIAN in action: assisted design of data science pipelines. Proc. VLDB Endow. 15(12), 3714\u20133717 (2022). https:\/\/doi.org\/10.14778\/3554821.3554882","journal-title":"Proc. VLDB Endow."},{"issue":"12","key":"835_CR43","doi-asserted-by":"publisher","first-page":"1954","DOI":"10.14778\/3352063.3352108","volume":"12","author":"EK Rezig","year":"2019","unstructured":"Rezig, E.K., Cao, L., Stonebraker, M., Simonini, G., Tao, W., Madden, S., Ouzzani, M., Tang, N., Elmagarmid, A.K.: Data civilizer 2.0: a holistic framework for data preparation and analytics. Proc. VLDB Endow. 12(12), 1954\u20131957 (2019)","journal-title":"Proc. VLDB Endow."},{"key":"835_CR44","unstructured":"Schelter, S., B\u00f6se, J.H., Kirschnick, J., Klein, T., Seufert, S., Amazon: declarative metadata management: a missing piece in end-to-end machine learning. SysML (2018). https:\/\/api.semanticscholar.org\/CorpusID:52841157"},{"issue":"3","key":"835_CR45","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/2480741.2480748","volume":"45","author":"F Serban","year":"2013","unstructured":"Serban, F., Vanschoren, J., Kietz, J.U., Bernstein, A.: A survey of intelligent assistants for data analysis. ACM Comput. Surv. CSUR 45(3), 1\u201335 (2013)","journal-title":"ACM Comput. Surv. CSUR"},{"issue":"1","key":"835_CR46","doi-asserted-by":"publisher","first-page":"148","DOI":"10.1109\/JPROC.2015.2494218","volume":"104","author":"B Shahriari","year":"2015","unstructured":"Shahriari, B., et al.: Taking the human out of the loop: a review of Bayesian optimization. Proc. IEEE 104(1), 148\u2013175 (2015)","journal-title":"Proc. IEEE"},{"key":"835_CR47","doi-asserted-by":"publisher","unstructured":"Shang, Z., et\u00a0al.: Democratizing data science through interactive curation of ml pipelines. In: SIGMOD\u201919, pp. 1171\u20131188. ACM (2019). https:\/\/doi.org\/10.1145\/3299869.3319863, https:\/\/doi.org\/10.1145\/3299869.3319863","DOI":"10.1145\/3299869.3319863"},{"key":"835_CR48","doi-asserted-by":"crossref","unstructured":"Smith, M.J., Sala, C., Kanter, J.M., Veeramachaneni, K.: The machine learning bazaar: harnessing the ml ecosystem for effective system development. In: Proceedings of the 2020 ACM SIGMOD International Conference on Management of Data, pp. 785\u2013800 (2020)","DOI":"10.1145\/3318464.3386146"},{"key":"835_CR49","unstructured":"Surowiecki, J.: The Wisdom of Crowds. Knopf Doubleday Publishing Group (2005). https:\/\/books.google.de\/books?id=hHUsHOHqVzEC"},{"key":"835_CR50","doi-asserted-by":"crossref","unstructured":"Vanschoren, J.: Meta-learning. In: Automated Machine Learning, pp. 35\u201361. Springer, Cham (2019)","DOI":"10.1007\/978-3-030-05318-5_2"},{"issue":"2","key":"835_CR51","doi-asserted-by":"publisher","first-page":"49","DOI":"10.1145\/2641190.2641198","volume":"15","author":"J Vanschoren","year":"2014","unstructured":"Vanschoren, J., Van Rijn, J.N., Bischl, B., Torgo, L.: Openml: networked science in machine learning. ACM SIGKDD Explor. Newsl 15(2), 49\u201360 (2014)","journal-title":"ACM SIGKDD Explor. Newsl"},{"key":"835_CR52","doi-asserted-by":"crossref","unstructured":"Vartak, M., Subramanyam, H., Lee, W.E., Viswanathan, S., Husnoo, S., Madden, S., Zaharia, M.: Modeldb: a system for machine learning model management. In: Proceedings of the Workshop on Human-In-the-Loop Data Analytics, pp. 1\u20133 (2016)","DOI":"10.1145\/2939502.2939516"},{"key":"835_CR53","doi-asserted-by":"crossref","unstructured":"Vlot, A., Maghsudi, S., Ohler, U.: Semitones: single-cell marker identification by enrichment scoring. Cold Spring Harbor Laboratory (2020)","DOI":"10.1101\/2020.11.17.386664"},{"key":"835_CR54","unstructured":"Wang, Y., Wang, L., Li, Y., He, D., Chen, W., Liu, T.: A theoretical analysis of ndcg ranking measures. In: Proceedings of the 26th Annual Conference on Learning Theory (COLT 2013), vol.\u00a08, p.\u00a06 (2013)"},{"key":"835_CR55","unstructured":"Yao, Q., et\u00a0al.: Taking human out of learning applications: a survey on automated machine learning. arXiv:1810.13306 (2018)"},{"issue":"4","key":"835_CR56","first-page":"39","volume":"41","author":"M Zaharia","year":"2018","unstructured":"Zaharia, M., Chen, A., Davidson, A., Ghodsi, A., Hong, S.A., Konwinski, A., Murching, S., Nykodym, T., Ogilvie, P., Parkhe, M., et al.: Accelerating the machine learning lifecycle with mlflow. IEEE Data Eng. Bull. 41(4), 39\u201345 (2018)","journal-title":"IEEE Data Eng. Bull."},{"key":"835_CR57","doi-asserted-by":"publisher","unstructured":"Zeng, Z., Tung, A.K.H., Wang, J., Feng, J., Zhou, L.: Comparing stars: on approximating graph edit distance. Proc. VLDB Endow. 2(1), 25\u201336 (2009). https:\/\/doi.org\/10.14778\/1687627.1687631","DOI":"10.14778\/1687627.1687631"}],"container-title":["The VLDB Journal"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s00778-024-00835-2.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s00778-024-00835-2\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s00778-024-00835-2.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,11,11]],"date-time":"2024-11-11T04:41:16Z","timestamp":1731300076000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s00778-024-00835-2"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,2,13]]},"references-count":57,"journal-issue":{"issue":"4","published-print":{"date-parts":[[2024,7]]}},"alternative-id":["835"],"URL":"https:\/\/doi.org\/10.1007\/s00778-024-00835-2","relation":{},"ISSN":["1066-8888","0949-877X"],"issn-type":[{"value":"1066-8888","type":"print"},{"value":"0949-877X","type":"electronic"}],"subject":[],"published":{"date-parts":[[2024,2,13]]},"assertion":[{"value":"31 January 2023","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"25 December 2023","order":2,"name":"revised","label":"Revised","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"2 January 2024","order":3,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"13 February 2024","order":4,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}}]}}