{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,2,20]],"date-time":"2026-02-20T08:37:56Z","timestamp":1771576676573,"version":"3.50.1"},"publisher-location":"New York, NY, USA","reference-count":58,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,5,23]],"date-time":"2022-05-23T00:00:00Z","timestamp":1653264000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"NSF","award":["1908762"],"award-info":[{"award-number":["1908762"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,5,23]]},"DOI":"10.1145\/3524842.3528458","type":"proceedings-article","created":{"date-parts":[[2022,10,18]],"date-time":"2022-10-18T00:08:36Z","timestamp":1666051716000},"page":"156-166","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":16,"title":["How to improve deep learning for software analytics"],"prefix":"10.1145","author":[{"given":"Rahul","family":"Yedida","sequence":"first","affiliation":[{"name":"NC State University"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Tim","family":"Menzies","sequence":"additional","affiliation":[{"name":"NC State University"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2022,10,17]]},"reference":[{"key":"e_1_3_2_1_1_1","volume-title":"How to\" DODGE\" Complex Software Analytics","author":"Agrawal Amritanshu","year":"2019","unstructured":"Amritanshu Agrawal , Wei Fu , Di Chen , Xipeng Shen , and Tim Menzies . 2019. How to\" DODGE\" Complex Software Analytics . IEEE Transactions on Software Engineering ( 2019 ). Amritanshu Agrawal, Wei Fu, Di Chen, Xipeng Shen, and Tim Menzies. 2019. How to\" DODGE\" Complex Software Analytics. IEEE Transactions on Software Engineering (2019)."},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1145\/3180155.3180197"},{"key":"e_1_3_2_1_3_1","volume-title":"How, When","author":"Agrawal Amritanshu","year":"2021","unstructured":"Amritanshu Agrawal , Xueqi Yang , Rishabh Agrawal , Rahul Yedida , Xipeng Shen , and Tim Menzies . 2021. Simpler Hyperparameter Optimization for Software Analytics: Why , How, When . IEEE Transactions on Software Engineering ( 2021 ). Amritanshu Agrawal, Xueqi Yang, Rishabh Agrawal, Rahul Yedida, Xipeng Shen, and Tim Menzies. 2021. Simpler Hyperparameter Optimization for Software Analytics: Why, How, When. IEEE Transactions on Software Engineering (2021)."},{"key":"e_1_3_2_1_4_1","volume-title":"code2seq: Generating sequences from structured representations of code. arXiv preprint arXiv:1808.01400","author":"Alon Uri","year":"2018","unstructured":"Uri Alon , Shaked Brody , Omer Levy , and Eran Yahav . 2018. code2seq: Generating sequences from structured representations of code. arXiv preprint arXiv:1808.01400 ( 2018 ). Uri Alon, Shaked Brody, Omer Levy, and Eran Yahav. 2018. code2seq: Generating sequences from structured representations of code. arXiv preprint arXiv:1808.01400 (2018)."},{"key":"e_1_3_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1145\/3290353"},{"key":"e_1_3_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.infsof.2018.12.009"},{"key":"e_1_3_2_1_7_1","volume-title":"Bad smells in code. Refactoring: Improving the design of existing code 1","author":"Beck Kent","year":"1999","unstructured":"Kent Beck , Martin Fowler , and Grandma Beck . 1999. Bad smells in code. Refactoring: Improving the design of existing code 1 , 1999 (1999), 75--88. Kent Beck, Martin Fowler, and Grandma Beck. 1999. Bad smells in code. Refactoring: Improving the design of existing code 1, 1999 (1999), 75--88."},{"key":"e_1_3_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.5555\/1622407.1622416"},{"key":"e_1_3_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1145\/3180155.3180240"},{"key":"e_1_3_2_1_10_1","volume-title":"Learning efficient object detection models with knowledge distillation. Advances in neural information processing systems 30","author":"Chen Guobin","year":"2017","unstructured":"Guobin Chen , Wongun Choi , Xiang Yu , Tony Han , and Manmohan Chandraker . 2017. Learning efficient object detection models with knowledge distillation. Advances in neural information processing systems 30 ( 2017 ). Guobin Chen, Wongun Choi, Xiang Yu, Tony Han, and Manmohan Chandraker. 2017. Learning efficient object detection models with knowledge distillation. Advances in neural information processing systems 30 (2017)."},{"key":"e_1_3_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00489"},{"key":"e_1_3_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10664-020-09898-5"},{"key":"e_1_3_2_1_13_1","volume-title":"G\u00e9rard Ben Arous, and Yann LeCun","author":"Choromanska Anna","year":"2015","unstructured":"Anna Choromanska , Mikael Henaff , Michael Mathieu , G\u00e9rard Ben Arous, and Yann LeCun . 2015 . The loss surfaces of multilayer networks. In Artificial intelligence and statistics. PMLR , 192--204. Anna Choromanska, Mikael Henaff, Michael Mathieu, G\u00e9rard Ben Arous, and Yann LeCun. 2015. The loss surfaces of multilayer networks. In Artificial intelligence and statistics. PMLR, 192--204."},{"key":"e_1_3_2_1_14_1","volume-title":"Approximation by superpositions of a sigmoidal function. Mathematics of control, signals and systems 2, 4","author":"Cybenko George","year":"1989","unstructured":"George Cybenko . 1989. Approximation by superpositions of a sigmoidal function. Mathematics of control, signals and systems 2, 4 ( 1989 ), 303--314. George Cybenko. 1989. Approximation by superpositions of a sigmoidal function. Mathematics of control, signals and systems 2, 4 (1989), 303--314."},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1016\/S0164-1212(03)00240-1"},{"key":"e_1_3_2_1_16_1","volume-title":"Gradient descent provably optimizes over-parameterized neural networks. arXiv preprint arXiv:1810.02054","author":"Du Simon S","year":"2018","unstructured":"Simon S Du , Xiyu Zhai , Barnabas Poczos , and Aarti Singh . 2018. Gradient descent provably optimizes over-parameterized neural networks. arXiv preprint arXiv:1810.02054 ( 2018 ). Simon S Du, Xiyu Zhai, Barnabas Poczos, and Aarti Singh. 2018. Gradient descent provably optimizes over-parameterized neural networks. arXiv preprint arXiv:1810.02054 (2018)."},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.5555\/2938006.2938019"},{"key":"e_1_3_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICSM.2013.56"},{"key":"e_1_3_2_1_19_1","volume-title":"Forget me not: A Gentle Reminder to Mind the Simple Multi-Layer Perceptron Baseline for Text Classification. arXiv preprint arXiv:2109.03777","author":"Galke Lukas","year":"2021","unstructured":"Lukas Galke and Ansgar Scherp . 2021. Forget me not: A Gentle Reminder to Mind the Simple Multi-Layer Perceptron Baseline for Text Classification. arXiv preprint arXiv:2109.03777 ( 2021 ). Lukas Galke and Ansgar Scherp. 2021. Forget me not: A Gentle Reminder to Mind the Simple Multi-Layer Perceptron Baseline for Text Classification. arXiv preprint arXiv:2109.03777 (2021)."},{"key":"e_1_3_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.1145\/3468264.3468553"},{"key":"e_1_3_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-021-01453-z"},{"key":"e_1_3_2_1_22_1","unstructured":"Melinda R Hess and Jeffrey D Kromrey. 2004. Robust confidence intervals for effect sizes: A comparative study of Cohen'sd and Cliff's delta under non-normality and heterogeneous variances. In annual meeting of the American Educational Research Association. Citeseer 1--30.  Melinda R Hess and Jeffrey D Kromrey. 2004. Robust confidence intervals for effect sizes: A comparative study of Cohen'sd and Cliff's delta under non-normality and heterogeneous variances. In annual meeting of the American Educational Research Association. Citeseer 1--30."},{"key":"e_1_3_2_1_23_1","volume-title":"Distilling the knowledge in a neural network. arXiv preprint arXiv:1503.02531","author":"Hinton Geoffrey","year":"2015","unstructured":"Geoffrey Hinton , Oriol Vinyals , and Jeff Dean . 2015. Distilling the knowledge in a neural network. arXiv preprint arXiv:1503.02531 ( 2015 ). Geoffrey Hinton, Oriol Vinyals, and Jeff Dean. 2015. Distilling the knowledge in a neural network. arXiv preprint arXiv:1503.02531 (2015)."},{"key":"e_1_3_2_1_24_1","volume-title":"Long short-term memory. Neural computation 9, 8","author":"Hochreiter Sepp","year":"1997","unstructured":"Sepp Hochreiter and J\u00fcrgen Schmidhuber . 1997. Long short-term memory. Neural computation 9, 8 ( 1997 ), 1735--1780. Sepp Hochreiter and J\u00fcrgen Schmidhuber. 1997. Long short-term memory. Neural computation 9, 8 (1997), 1735--1780."},{"key":"e_1_3_2_1_25_1","volume-title":"Multilayer feedforward networks are universal approximators. Neural networks 2, 5","author":"Hornik Kurt","year":"1989","unstructured":"Kurt Hornik , Maxwell Stinchcombe , and Halbert White . 1989. Multilayer feedforward networks are universal approximators. Neural networks 2, 5 ( 1989 ), 359--366. Kurt Hornik, Maxwell Stinchcombe, and Halbert White. 1989. Multilayer feedforward networks are universal approximators. Neural networks 2, 5 (1989), 359--366."},{"key":"e_1_3_2_1_26_1","volume-title":"Neural tangent kernel: Convergence and generalization in neural networks. arXiv preprint arXiv:1806.07572","author":"Jacot Arthur","year":"2018","unstructured":"Arthur Jacot , Franck Gabriel , and Cl\u00e9ment Hongler . 2018. Neural tangent kernel: Convergence and generalization in neural networks. arXiv preprint arXiv:1806.07572 ( 2018 ). Arthur Jacot, Franck Gabriel, and Cl\u00e9ment Hongler. 2018. Neural tangent kernel: Convergence and generalization in neural networks. arXiv preprint arXiv:1806.07572 (2018)."},{"key":"e_1_3_2_1_27_1","volume-title":"Sequence-level knowledge distillation. arXiv preprint arXiv:1606.07947","author":"Kim Yoon","year":"2016","unstructured":"Yoon Kim and Alexander M Rush . 2016. Sequence-level knowledge distillation. arXiv preprint arXiv:1606.07947 ( 2016 ). Yoon Kim and Alexander M Rush. 2016. Sequence-level knowledge distillation. arXiv preprint arXiv:1606.07947 (2016)."},{"key":"e_1_3_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.jss.2020.110783"},{"key":"e_1_3_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1109\/TSE.2019.2936376"},{"key":"e_1_3_2_1_30_1","volume-title":"2018 IEEE\/ACM 15th International Conference on Mining Software Repositories (MSR). IEEE, 554--563","author":"Menzies Tim","year":"2018","unstructured":"Tim Menzies , Suvodeep Majumder , Nikhila Balaji , Katie Brey , and Wei Fu . 2018 . 500+ times faster than deep learning:(a case study exploring faster methods for text mining stackoverflow) . In 2018 IEEE\/ACM 15th International Conference on Mining Software Repositories (MSR). IEEE, 554--563 . Tim Menzies, Suvodeep Majumder, Nikhila Balaji, Katie Brey, and Wei Fu. 2018. 500+ times faster than deep learning:(a case study exploring faster methods for text mining stackoverflow). In 2018 IEEE\/ACM 15th International Conference on Mining Software Repositories (MSR). IEEE, 554--563."},{"key":"e_1_3_2_1_31_1","volume-title":"Efficient estimation of word representations in vector space. arXiv preprint arXiv:1301.3781","author":"Mikolov Tomas","year":"2013","unstructured":"Tomas Mikolov , Kai Chen , Greg Corrado , and Jeffrey Dean . 2013. Efficient estimation of word representations in vector space. arXiv preprint arXiv:1301.3781 ( 2013 ). Tomas Mikolov, Kai Chen, Greg Corrado, and Jeffrey Dean. 2013. Efficient estimation of word representations in vector space. arXiv preprint arXiv:1301.3781 (2013)."},{"key":"e_1_3_2_1_32_1","unstructured":"Tomas Mikolov Ilya Sutskever Kai Chen Greg S Corrado and Jeff Dean. 2013. Distributed representations of words and phrases and their compositionality. In Advances in neural information processing systems. 3111--3119.  Tomas Mikolov Ilya Sutskever Kai Chen Greg S Corrado and Jeff Dean. 2013. Distributed representations of words and phrases and their compositionality. In Advances in neural information processing systems. 3111--3119."},{"key":"e_1_3_2_1_33_1","volume-title":"On the number of linear regions of deep neural networks. arXiv preprint arXiv:1402.1869","author":"Mont\u00fafar Guido","year":"2014","unstructured":"Guido Mont\u00fafar , Razvan Pascanu , Kyunghyun Cho , and Yoshua Bengio . 2014. On the number of linear regions of deep neural networks. arXiv preprint arXiv:1402.1869 ( 2014 ). Guido Mont\u00fafar, Razvan Pascanu, Kyunghyun Cho, and Yoshua Bengio. 2014. On the number of linear regions of deep neural networks. arXiv preprint arXiv:1402.1869 (2014)."},{"key":"e_1_3_2_1_34_1","volume-title":"IFIP Central and East European Conference on Software Engineering Techniques. Springer, 252--266","author":"Moser Raimund","year":"2007","unstructured":"Raimund Moser , Pekka Abrahamsson , Witold Pedrycz , Alberto Sillitti , and Giancarlo Succi . 2007 . A case study on the impact of refactoring on quality and productivity in an agile team . In IFIP Central and East European Conference on Software Engineering Techniques. Springer, 252--266 . Raimund Moser, Pekka Abrahamsson, Witold Pedrycz, Alberto Sillitti, and Giancarlo Succi. 2007. A case study on the impact of refactoring on quality and productivity in an agile team. In IFIP Central and East European Conference on Software Engineering Techniques. Springer, 252--266."},{"key":"e_1_3_2_1_35_1","unstructured":"Vinod Nair and Geoffrey E Hinton. 2010. Rectified linear units improve restricted boltzmann machines. In Icml.  Vinod Nair and Geoffrey E Hinton. 2010. Rectified linear units improve restricted boltzmann machines. In Icml."},{"key":"e_1_3_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICSE.2015.244"},{"key":"e_1_3_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.1109\/IJCNN.2016.7727212"},{"key":"e_1_3_2_1_38_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00409"},{"key":"e_1_3_2_1_39_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.jss.2020.110693"},{"key":"e_1_3_2_1_40_1","volume-title":"Integrating Tree Path in Transformer for Code Representation. Advances in Neural Information Processing Systems 34","author":"Peng Han","year":"2021","unstructured":"Han Peng , Ge Li , Wenhan Wang , Yunfei Zhao , and Zhi Jin . 2021. Integrating Tree Path in Transformer for Code Representation. Advances in Neural Information Processing Systems 34 ( 2021 ). Han Peng, Ge Li, Wenhan Wang, Yunfei Zhao, and Zhi Jin. 2021. Integrating Tree Path in Transformer for Code Representation. Advances in Neural Information Processing Systems 34 (2021)."},{"key":"e_1_3_2_1_41_1","volume-title":"International Conference on Machine Learning. PMLR, 5142--5151","author":"Phuong Mary","year":"2019","unstructured":"Mary Phuong and Christoph Lampert . 2019 . Towards understanding knowledge distillation . In International Conference on Machine Learning. PMLR, 5142--5151 . Mary Phuong and Christoph Lampert. 2019. Towards understanding knowledge distillation. In International Conference on Machine Learning. PMLR, 5142--5151."},{"key":"e_1_3_2_1_42_1","volume-title":"Making the most of small Software Engineering datasets with modern machine learning","author":"Aron Prenner Julian Aron","year":"2021","unstructured":"Julian Aron Aron Prenner and Romain Robbes . 2021. Making the most of small Software Engineering datasets with modern machine learning . IEEE Transactions on Software Engineering ( 2021 ). Julian Aron Aron Prenner and Romain Robbes. 2021. Making the most of small Software Engineering datasets with modern machine learning. IEEE Transactions on Software Engineering (2021)."},{"key":"e_1_3_2_1_43_1","unstructured":"Alec Radford Jeffrey Wu Rewon Child David Luan Dario Amodei Ilya Sutskever etal 2019. Language models are unsupervised multitask learners. OpenAI blog 1 8 (2019) 9.  Alec Radford Jeffrey Wu Rewon Child David Luan Dario Amodei Ilya Sutskever et al. 2019. Language models are unsupervised multitask learners. OpenAI blog 1 8 (2019) 9."},{"key":"e_1_3_2_1_44_1","volume-title":"Learning representations by back-propagating errors. nature 323, 6088","author":"Rumelhart David E","year":"1986","unstructured":"David E Rumelhart , Geoffrey E Hinton , and Ronald J Williams . 1986. Learning representations by back-propagating errors. nature 323, 6088 ( 1986 ), 533--536. David E Rumelhart, Geoffrey E Hinton, and Ronald J Williams. 1986. Learning representations by back-propagating errors. nature 323, 6088 (1986), 533--536."},{"key":"e_1_3_2_1_45_1","doi-asserted-by":"publisher","DOI":"10.1145\/2675067"},{"key":"e_1_3_2_1_46_1","volume-title":"Proceedings of the 32nd international conference on neural information processing systems. 2488--2498","author":"Santurkar Shibani","year":"2018","unstructured":"Shibani Santurkar , Dimitris Tsipras , Andrew Ilyas , and Aleksander M\u0105dry . 2018 . How does batch normalization help optimization? . In Proceedings of the 32nd international conference on neural information processing systems. 2488--2498 . Shibani Santurkar, Dimitris Tsipras, Andrew Ilyas, and Aleksander M\u0105dry. 2018. How does batch normalization help optimization?. In Proceedings of the 32nd international conference on neural information processing systems. 2488--2498."},{"key":"e_1_3_2_1_47_1","doi-asserted-by":"publisher","DOI":"10.1145\/1852786.1852797"},{"key":"e_1_3_2_1_48_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.neuroimage.2014.06.077"},{"key":"e_1_3_2_1_49_1","doi-asserted-by":"publisher","DOI":"10.1109\/APSEC.2010.46"},{"key":"e_1_3_2_1_51_1","doi-asserted-by":"publisher","DOI":"10.1145\/2601248.2601268"},{"key":"e_1_3_2_1_52_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.infsof.2013.08.002"},{"key":"e_1_3_2_1_53_1","volume-title":"Richard Kinh Gian Do, and Kaori Togashi","author":"Yamashita Rikiya","year":"2018","unstructured":"Rikiya Yamashita , Mizuho Nishio , Richard Kinh Gian Do, and Kaori Togashi . 2018 . Convolutional neural networks: an overview and application in radiology. Insights into imaging 9, 4 (2018), 611--629. Rikiya Yamashita, Mizuho Nishio, Richard Kinh Gian Do, and Kaori Togashi. 2018. Convolutional neural networks: an overview and application in radiology. Insights into imaging 9, 4 (2018), 611--629."},{"key":"e_1_3_2_1_54_1","volume-title":"On the Value of Oversampling for Deep Learning in Software Defect Prediction","author":"Yedida Rahul","year":"2021","unstructured":"Rahul Yedida and Tim Menzies . 2021. On the Value of Oversampling for Deep Learning in Software Defect Prediction . IEEE Transactions on Software Engineering ( 2021 ). Rahul Yedida and Tim Menzies. 2021. On the Value of Oversampling for Deep Learning in Software Defect Prediction. IEEE Transactions on Software Engineering (2021)."},{"key":"e_1_3_2_1_55_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.754"},{"key":"e_1_3_2_1_56_1","doi-asserted-by":"publisher","DOI":"10.1145\/1985362.1985366"},{"key":"e_1_3_2_1_57_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-10590-1_53"},{"key":"e_1_3_2_1_58_1","volume-title":"Software Vulnerability Detection via Deep Learning over Disaggregated Code Graph Representation. arXiv preprint arXiv:2109.03341","author":"Zhuang Yufan","year":"2021","unstructured":"Yufan Zhuang , Sahil Suneja , Veronika Thost , Giacomo Domeniconi , Alessandro Morari , and Jim Laredo . 2021. Software Vulnerability Detection via Deep Learning over Disaggregated Code Graph Representation. arXiv preprint arXiv:2109.03341 ( 2021 ). Yufan Zhuang, Sahil Suneja, Veronika Thost, Giacomo Domeniconi, Alessandro Morari, and Jim Laredo. 2021. Software Vulnerability Detection via Deep Learning over Disaggregated Code Graph Representation. arXiv preprint arXiv:2109.03341 (2021)."},{"key":"e_1_3_2_1_59_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10994-019-05839-6"}],"event":{"name":"MSR '22: 19th International Conference on Mining Software Repositories","location":"Pittsburgh Pennsylvania","acronym":"MSR '22","sponsor":["SIGSOFT ACM Special Interest Group on Software Engineering","IEEE CS"]},"container-title":["Proceedings of the 19th International Conference on Mining Software Repositories"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3524842.3528458","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3524842.3528458","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3524842.3528458","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T18:09:35Z","timestamp":1750183775000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3524842.3528458"}},"subtitle":["(a case study with code smell detection)"],"short-title":[],"issued":{"date-parts":[[2022,5,23]]},"references-count":58,"alternative-id":["10.1145\/3524842.3528458","10.1145\/3524842"],"URL":"https:\/\/doi.org\/10.1145\/3524842.3528458","relation":{},"subject":[],"published":{"date-parts":[[2022,5,23]]},"assertion":[{"value":"2022-10-17","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}