{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,22]],"date-time":"2026-07-22T04:48:59Z","timestamp":1784695739917,"version":"3.55.0"},"publisher-location":"New York, NY, USA","reference-count":68,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,11,7]],"date-time":"2022-11-07T00:00:00Z","timestamp":1667779200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/100000001","name":"National Science Foundation","doi-asserted-by":"publisher","award":["CCF-2107405, SHF-1845893, SHF-1414172, SHF-2107592, IIS-2221943"],"award-info":[{"award-number":["CCF-2107405, SHF-1845893, SHF-1414172, SHF-2107592, IIS-2221943"]}],"id":[{"id":"10.13039\/100000001","id-type":"DOI","asserted-by":"publisher"}]},{"name":"IBM","award":["Faculty Award"],"award-info":[{"award-number":["Faculty Award"]}]},{"name":"UC Davis College of Engineering","award":["Dean?s Distinguished Fellowship"],"award-info":[{"award-number":["Dean?s Distinguished Fellowship"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,11,7]]},"DOI":"10.1145\/3540250.3549162","type":"proceedings-article","created":{"date-parts":[[2022,11,9]],"date-time":"2022-11-09T20:46:22Z","timestamp":1668026782000},"page":"18-30","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":97,"title":["NatGen: generative pre-training by \u201cnaturalizing\u201d source code"],"prefix":"10.1145","author":[{"given":"Saikat","family":"Chakraborty","sequence":"first","affiliation":[{"name":"Columbia University, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Toufique","family":"Ahmed","sequence":"additional","affiliation":[{"name":"University of California at Davis, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yangruibo","family":"Ding","sequence":"additional","affiliation":[{"name":"Columbia University, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Premkumar T.","family":"Devanbu","sequence":"additional","affiliation":[{"name":"University of California at Davis, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Baishakhi","family":"Ray","sequence":"additional","affiliation":[{"name":"Columbia University, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2022,11,9]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.5281\/zenodo.6977595"},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2020.acl-main.449"},{"key":"e_1_3_2_1_3_1","volume-title":"Unified Pre-training for Program Understanding and Generation. In 2021 Annual Conference of the North American Chapter of the Association for Computational Linguistics (NAACL).","author":"Ahmad Wasi Uddin","year":"2021","unstructured":"Wasi Uddin Ahmad , Saikat Chakraborty , Baishakhi Ray , and Kai-Wei Chang . 2021 . Unified Pre-training for Program Understanding and Generation. In 2021 Annual Conference of the North American Chapter of the Association for Computational Linguistics (NAACL). Wasi Uddin Ahmad, Saikat Chakraborty, Baishakhi Ray, and Kai-Wei Chang. 2021. Unified Pre-training for Program Understanding and Generation. In 2021 Annual Conference of the North American Chapter of the Association for Computational Linguistics (NAACL)."},{"key":"e_1_3_2_1_4_1","volume-title":"Saikat Chakraborty, and Kai-Wei Chang.","author":"Ahmad Wasi Uddin","year":"2021","unstructured":"Wasi Uddin Ahmad , Md Golam Rahman Tushar , Saikat Chakraborty, and Kai-Wei Chang. 2021 . AVATAR : A Parallel Corpus for Java-Python Program Translation . arxiv:2108.11590. Wasi Uddin Ahmad, Md Golam Rahman Tushar, Saikat Chakraborty, and Kai-Wei Chang. 2021. AVATAR: A Parallel Corpus for Java-Python Program Translation. arxiv:2108.11590."},{"key":"e_1_3_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1145\/3359591.3359735"},{"key":"e_1_3_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1145\/2786805.2786849"},{"key":"e_1_3_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1145\/3212695"},{"key":"e_1_3_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1145\/3385412.3385997"},{"key":"e_1_3_2_1_9_1","unstructured":"Miltiadis Allamanis Marc Brockschmidt and Mahmoud Khademi. 2017. Learning to represent programs with graphs. arXiv preprint arXiv:1711.00740. \t\t\t\t  Miltiadis Allamanis Marc Brockschmidt and Mahmoud Khademi. 2017. Learning to represent programs with graphs. arXiv preprint arXiv:1711.00740."},{"key":"e_1_3_2_1_10_1","unstructured":"Matthew Amodio Swarat Chaudhuri and Thomas W Reps. 2017. Neural attribute machines for program generation. arXiv preprint arXiv:1705.09231. \t\t\t\t  Matthew Amodio Swarat Chaudhuri and Thomas W Reps. 2017. Neural attribute machines for program generation. arXiv preprint arXiv:1705.09231."},{"key":"e_1_3_2_1_11_1","unstructured":"Tom B. Brown Benjamin Mann Nick Ryder Melanie Subbiah Jared Kaplan Prafulla Dhariwal Arvind Neelakantan Pranav Shyam Girish Sastry Amanda Askell Sandhini Agarwal Ariel Herbert-Voss Gretchen Krueger Tom Henighan Rewon Child Aditya Ramesh Daniel M. Ziegler Jeffrey Wu Clemens Winter Christopher Hesse Mark Chen Eric Sigler Mateusz Litwin Scott Gray Benjamin Chess Jack Clark Christopher Berner Sam McCandlish Alec Radford Ilya Sutskever and Dario Amodei. 2020. Language Models are Few-Shot Learners. arxiv:2005.14165. \t\t\t\t  Tom B. Brown Benjamin Mann Nick Ryder Melanie Subbiah Jared Kaplan Prafulla Dhariwal Arvind Neelakantan Pranav Shyam Girish Sastry Amanda Askell Sandhini Agarwal Ariel Herbert-Voss Gretchen Krueger Tom Henighan Rewon Child Aditya Ramesh Daniel M. Ziegler Jeffrey Wu Clemens Winter Christopher Hesse Mark Chen Eric Sigler Mateusz Litwin Scott Gray Benjamin Chess Jack Clark Christopher Berner Sam McCandlish Alec Radford Ilya Sutskever and Dario Amodei. 2020. Language Models are Few-Shot Learners. arxiv:2005.14165."},{"key":"e_1_3_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1145\/3377816.3381720"},{"key":"e_1_3_2_1_13_1","volume-title":"Do programmers prefer predictable expressions in code? Cognitive science, 44, 12","author":"Casalnuovo Casey","year":"2020","unstructured":"Casey Casalnuovo , Kevin Lee , Hulin Wang , Prem Devanbu , and Emily Morgan . 2020. Do programmers prefer predictable expressions in code? Cognitive science, 44, 12 ( 2020 ), e12921. Casey Casalnuovo, Kevin Lee, Hulin Wang, Prem Devanbu, and Emily Morgan. 2020. Do programmers prefer predictable expressions in code? Cognitive science, 44, 12 (2020), e12921."},{"key":"e_1_3_2_1_14_1","volume-title":"Proceedings of the 42nd Annual Meeting of the Cognitive Science Society.","author":"Casalnuovo Casey","year":"2020","unstructured":"Casey Casalnuovo , E Morgan , and P Devanbu . 2020 . Does surprisal predict code comprehension difficulty . In Proceedings of the 42nd Annual Meeting of the Cognitive Science Society. Casey Casalnuovo, E Morgan, and P Devanbu. 2020. Does surprisal predict code comprehension difficulty. In Proceedings of the 42nd Annual Meeting of the Cognitive Science Society."},{"key":"e_1_3_2_1_15_1","first-page":"1","article-title":"CODIT: Code Editing with Tree-Based Neural Models","volume":"1","author":"Chakraborty Saikat","year":"2020","unstructured":"Saikat Chakraborty , Yangruibo Ding , Miltiadis Allamanis , and Baishakhi Ray . 2020 . CODIT: Code Editing with Tree-Based Neural Models . IEEE Transactions on Software Engineering , 1 (2020), 1 \u2013 1 . Saikat Chakraborty, Yangruibo Ding, Miltiadis Allamanis, and Baishakhi Ray. 2020. CODIT: Code Editing with Tree-Based Neural Models. IEEE Transactions on Software Engineering, 1 (2020), 1\u20131.","journal-title":"IEEE Transactions on Software Engineering"},{"key":"e_1_3_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/ASE51524.2021.9678559"},{"key":"e_1_3_2_1_17_1","unstructured":"Mark Chen Jerry Tworek Heewoo Jun Qiming Yuan Henrique Ponde de Oliveira Pinto Jared Kaplan Harri Edwards Yuri Burda Nicholas Joseph Greg Brockman Alex Ray Raul Puri Gretchen Krueger Michael Petrov Heidy Khlaaf Girish Sastry Pamela Mishkin Brooke Chan Scott Gray Nick Ryder Mikhail Pavlov Alethea Power Lukasz Kaiser Mohammad Bavarian Clemens Winter Philippe Tillet Felipe Petroski Such Dave Cummings Matthias Plappert Fotios Chantzis Elizabeth Barnes Ariel Herbert-Voss William Hebgen Guss Alex Nichol Alex Paino Nikolas Tezak Jie Tang Igor Babuschkin Suchir Balaji Shantanu Jain William Saunders Christopher Hesse Andrew N. Carr Jan Leike Josh Achiam Vedant Misra Evan Morikawa Alec Radford Matthew Knight Miles Brundage Mira Murati Katie Mayer Peter Welinder Bob McGrew Dario Amodei Sam McCandlish Ilya Sutskever and Wojciech Zaremba. 2021. Evaluating Large Language Models Trained on Code. arxiv:2107.03374. \t\t\t\t  Mark Chen Jerry Tworek Heewoo Jun Qiming Yuan Henrique Ponde de Oliveira Pinto Jared Kaplan Harri Edwards Yuri Burda Nicholas Joseph Greg Brockman Alex Ray Raul Puri Gretchen Krueger Michael Petrov Heidy Khlaaf Girish Sastry Pamela Mishkin Brooke Chan Scott Gray Nick Ryder Mikhail Pavlov Alethea Power Lukasz Kaiser Mohammad Bavarian Clemens Winter Philippe Tillet Felipe Petroski Such Dave Cummings Matthias Plappert Fotios Chantzis Elizabeth Barnes Ariel Herbert-Voss William Hebgen Guss Alex Nichol Alex Paino Nikolas Tezak Jie Tang Igor Babuschkin Suchir Balaji Shantanu Jain William Saunders Christopher Hesse Andrew N. Carr Jan Leike Josh Achiam Vedant Misra Evan Morikawa Alec Radford Matthew Knight Miles Brundage Mira Murati Katie Mayer Peter Welinder Bob McGrew Dario Amodei Sam McCandlish Ilya Sutskever and Wojciech Zaremba. 2021. Evaluating Large Language Models Trained on Code. arxiv:2107.03374."},{"key":"e_1_3_2_1_18_1","volume-title":"International Conference on Learning Representations. https:\/\/openreview.net\/pdf?id=r1xMH1BtvB","author":"Clark Kevin","unstructured":"Kevin Clark , Minh-Thang Luong , Quoc V. Le , and Christopher D. Manning . 2020. ELECTRA: Pre-training Text Encoders as Discriminators Rather Than Generators . In International Conference on Learning Representations. https:\/\/openreview.net\/pdf?id=r1xMH1BtvB Kevin Clark, Minh-Thang Luong, Quoc V. Le, and Christopher D. Manning. 2020. ELECTRA: Pre-training Text Encoders as Discriminators Rather Than Generators. In International Conference on Learning Representations. https:\/\/openreview.net\/pdf?id=r1xMH1BtvB"},{"key":"e_1_3_2_1_19_1","volume-title":"Proceedings of the 2019 Conference of the North American","author":"Devlin Jacob","unstructured":"Jacob Devlin , Ming-Wei Chang , Kenton Lee , and Kristina Toutanova . [n. d.]. BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding . In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers) . Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. [n. d.]. BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers)."},{"key":"e_1_3_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2022.acl-long.436"},{"key":"e_1_3_2_1_21_1","volume-title":"Proceedings of the thirteenth international conference on artificial intelligence and statistics. 201\u2013208","author":"Erhan Dumitru","year":"2010","unstructured":"Dumitru Erhan , Aaron Courville , Yoshua Bengio , and Pascal Vincent . 2010 . Why does unsupervised pre-training help deep learning? In Proceedings of the thirteenth international conference on artificial intelligence and statistics. 201\u2013208 . Dumitru Erhan, Aaron Courville, Yoshua Bengio, and Pascal Vincent. 2010. Why does unsupervised pre-training help deep learning? In Proceedings of the thirteenth international conference on artificial intelligence and statistics. 201\u2013208."},{"key":"e_1_3_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2020.findings-emnlp.139"},{"key":"e_1_3_2_1_23_1","volume-title":"Refactoring: improving the design of existing code","author":"Fowler Martin","unstructured":"Martin Fowler . 2018. Refactoring: improving the design of existing code . Addison-Wesley Professional . Martin Fowler. 2018. Refactoring: improving the design of existing code. Addison-Wesley Professional."},{"key":"e_1_3_2_1_24_1","unstructured":"GitHub. 2022. GitHub Copilot (. https:\/\/copilot.github.com\/ \t\t\t\t  GitHub. 2022. GitHub Copilot (. https:\/\/copilot.github.com\/"},{"key":"e_1_3_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1145\/3368089.3409714"},{"key":"e_1_3_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1145\/3196398.3196432"},{"key":"e_1_3_2_1_27_1","volume-title":"Baselining & Evaluation. In 2020 35th IEEE\/ACM International Conference on Automated Software Engineering (ASE). 746\u2013757","author":"Gros David","year":"2020","unstructured":"David Gros , Hariharan Sezhiyan , Prem Devanbu , and Zhou Yu . 2020 . Code to Comment \u201cTranslation\u201d: Data, Metrics , Baselining & Evaluation. In 2020 35th IEEE\/ACM International Conference on Automated Software Engineering (ASE). 746\u2013757 . David Gros, Hariharan Sezhiyan, Prem Devanbu, and Zhou Yu. 2020. Code to Comment \u201cTranslation\u201d: Data, Metrics, Baselining & Evaluation. In 2020 35th IEEE\/ACM International Conference on Automated Software Engineering (ASE). 746\u2013757."},{"key":"e_1_3_2_1_28_1","volume-title":"GraphCodeBERT: Pre-training Code Representations with Data Flow. In International Conference on Learning Representations.","author":"Guo Daya","year":"2021","unstructured":"Daya Guo , Shuo Ren , Shuai Lu , Zhangyin Feng , Duyu Tang , Shujie Liu , Long Zhou , Nan Duan , Jian Yin , and Daxin Jiang . 2021 . GraphCodeBERT: Pre-training Code Representations with Data Flow. In International Conference on Learning Representations. Daya Guo, Shuo Ren, Shuai Lu, Zhangyin Feng, Duyu Tang, Shujie Liu, Long Zhou, Nan Duan, Jian Yin, and Daxin Jiang. 2021. GraphCodeBERT: Pre-training Code Representations with Data Flow. In International Conference on Learning Representations."},{"key":"e_1_3_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/P19-1082"},{"key":"e_1_3_2_1_30_1","doi-asserted-by":"crossref","unstructured":"Rahul Gupta Soham Pal Aditya Kanade and Shirish Shevade. 2017. DeepFix: Fixing Common C Language Errors by Deep Learning.. In AAAI. 1345\u20131351. \t\t\t\t  Rahul Gupta Soham Pal Aditya Kanade and Shirish Shevade. 2017. DeepFix: Fixing Common C Language Errors by Deep Learning.. In AAAI. 1345\u20131351.","DOI":"10.1609\/aaai.v31i1.10742"},{"key":"e_1_3_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.1145\/3106237.3106290"},{"key":"e_1_3_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1145\/2902362"},{"key":"e_1_3_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICSE.2012.6227135"},{"key":"e_1_3_2_1_34_1","unstructured":"Hamel Husain Ho-Hsiang Wu Tiferet Gazit Miltiadis Allamanis and Marc Brockschmidt. 2019. Codesearchnet challenge: Evaluating the state of semantic code search. arXiv preprint arXiv:1909.09436 arxiv:1909.09436 \t\t\t\t  Hamel Husain Ho-Hsiang Wu Tiferet Gazit Miltiadis Allamanis and Marc Brockschmidt. 2019. Codesearchnet challenge: Evaluating the state of semantic code search. arXiv preprint arXiv:1909.09436 arxiv:1909.09436"},{"key":"e_1_3_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/P16-1195"},{"key":"e_1_3_2_1_36_1","unstructured":"Srinivasan Iyer Ioannis Konstas Alvin Cheung and Luke Zettlemoyer. 2018. Mapping language to code in programmatic context. arXiv preprint arXiv:1808.09588. \t\t\t\t  Srinivasan Iyer Ioannis Konstas Alvin Cheung and Luke Zettlemoyer. 2018. Mapping language to code in programmatic context. arXiv preprint arXiv:1808.09588."},{"key":"e_1_3_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.1073\/pnas.1611835114"},{"key":"e_1_3_2_1_38_1","unstructured":"Marie-Anne Lachaux Baptiste Roziere Lowik Chanussot and Guillaume Lample. 2020. Unsupervised Translation of Programming Languages. arXiv preprint arXiv:2006.03511. \t\t\t\t  Marie-Anne Lachaux Baptiste Roziere Lowik Chanussot and Guillaume Lample. 2020. Unsupervised Translation of Programming Languages. arXiv preprint arXiv:2006.03511."},{"key":"e_1_3_2_1_39_1","volume-title":"Bart: Denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension. arXiv preprint arXiv:1910.13461.","author":"Lewis Mike","year":"2019","unstructured":"Mike Lewis , Yinhan Liu , Naman Goyal , Marjan Ghazvininejad , Abdelrahman Mohamed , Omer Levy , Ves Stoyanov , and Luke Zettlemoyer . 2019 . Bart: Denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension. arXiv preprint arXiv:1910.13461. Mike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad, Abdelrahman Mohamed, Omer Levy, Ves Stoyanov, and Luke Zettlemoyer. 2019. Bart: Denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension. arXiv preprint arXiv:1910.13461."},{"key":"e_1_3_2_1_40_1","unstructured":"Shangqing Liu Yu Chen Xiaofei Xie Jingkai Siow and Yang Liu. 2020. Retrieval-augmented generation for code summarization via hybrid gnn. arXiv preprint arXiv:2006.05405. \t\t\t\t  Shangqing Liu Yu Chen Xiaofei Xie Jingkai Siow and Yang Liu. 2020. Retrieval-augmented generation for code summarization via hybrid gnn. arXiv preprint arXiv:2006.05405."},{"key":"e_1_3_2_1_41_1","unstructured":"Yinhan Liu Myle Ott Naman Goyal Jingfei Du Mandar Joshi Danqi Chen Omer Levy Mike Lewis Luke Zettlemoyer and Veselin Stoyanov. 2019. RoBERTa: A Robustly Optimized BERT Pretraining Approach. arXiv preprint arXiv:1907.11692 arxiv:1907.11692 \t\t\t\t  Yinhan Liu Myle Ott Naman Goyal Jingfei Du Mandar Joshi Danqi Chen Omer Levy Mike Lewis Luke Zettlemoyer and Veselin Stoyanov. 2019. RoBERTa: A Robustly Optimized BERT Pretraining Approach. arXiv preprint arXiv:1907.11692 arxiv:1907.11692"},{"key":"e_1_3_2_1_42_1","unstructured":"Shuai Lu Daya Guo Shuo Ren Junjie Huang Alexey Svyatkovskiy Ambrosio Blanco Colin Clement Dawn Drain Daxin Jiang and Duyu Tang. 2021. CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation. arXiv preprint arXiv:2102.04664 arxiv:2102.04664 \t\t\t\t  Shuai Lu Daya Guo Shuo Ren Junjie Huang Alexey Svyatkovskiy Ambrosio Blanco Colin Clement Dawn Drain Daxin Jiang and Duyu Tang. 2021. CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation. arXiv preprint arXiv:2102.04664 arxiv:2102.04664"},{"key":"e_1_3_2_1_43_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICSE43902.2021.00041"},{"key":"e_1_3_2_1_44_1","unstructured":"Microsoft. 2021. CodeBLEU calculator (. https:\/\/github.com\/microsoft\/CodeXGLUE\/tree\/main\/Code-Code\/code-to-code-trans\/evaluator\/CodeBLEU \t\t\t\t  Microsoft. 2021. CodeBLEU calculator (. https:\/\/github.com\/microsoft\/CodeXGLUE\/tree\/main\/Code-Code\/code-to-code-trans\/evaluator\/CodeBLEU"},{"key":"e_1_3_2_1_45_1","unstructured":"Microsoft. 2022. CodeXGLUE Leaderboard (. https:\/\/microsoft.github.io\/CodeXGLUE\/ \t\t\t\t  Microsoft. 2022. CodeXGLUE Leaderboard (. https:\/\/microsoft.github.io\/CodeXGLUE\/"},{"key":"e_1_3_2_1_46_1","unstructured":"Changan Niu Chuanyi Li Vincent Ng Jidong Ge Liguo Huang and Bin Luo. 2022. SPT-Code: Sequence-to-Sequence Pre-Training for Learning the Representation of Source Code. arXiv preprint arXiv:2201.01549. \t\t\t\t  Changan Niu Chuanyi Li Vincent Ng Jidong Ge Liguo Huang and Bin Luo. 2022. SPT-Code: Sequence-to-Sequence Pre-Training for Learning the Representation of Source Code. arXiv preprint arXiv:2201.01549."},{"key":"e_1_3_2_1_47_1","volume-title":"Saikat Chakraborty, Baishakhi Ray, and Kai-Wei Chang.","author":"Parvez Md Rizwan","year":"2021","unstructured":"Md Rizwan Parvez , Wasi Uddin Ahmad , Saikat Chakraborty, Baishakhi Ray, and Kai-Wei Chang. 2021 . Retrieval Augmented Code Generation and Summarization . arXiv preprint arXiv:2108.11601. Md Rizwan Parvez, Wasi Uddin Ahmad, Saikat Chakraborty, Baishakhi Ray, and Kai-Wei Chang. 2021. Retrieval Augmented Code Generation and Summarization. arXiv preprint arXiv:2108.11601."},{"key":"e_1_3_2_1_48_1","doi-asserted-by":"publisher","DOI":"10.1145\/3468264.3468623"},{"key":"e_1_3_2_1_49_1","doi-asserted-by":"crossref","unstructured":"Long Phan Hieu Tran Daniel Le Hieu Nguyen James Anibal Alec Peltekian and Yanfang Ye. 2021. CoTexT: Multi-task Learning with Code-Text Transformer. arXiv preprint arXiv:2105.08645. \t\t\t\t  Long Phan Hieu Tran Daniel Le Hieu Nguyen James Anibal Alec Peltekian and Yanfang Ye. 2021. CoTexT: Multi-task Learning with Code-Text Transformer. arXiv preprint arXiv:2105.08645.","DOI":"10.18653\/v1\/2021.nlp4prog-1.5"},{"key":"e_1_3_2_1_50_1","doi-asserted-by":"publisher","DOI":"10.1145\/3460348"},{"key":"e_1_3_2_1_51_1","volume-title":"Language models are unsupervised multitask learners. OpenAI blog, 1, 8","author":"Radford Alec","year":"2019","unstructured":"Alec Radford , Jeffrey Wu , Rewon Child , David Luan , Dario Amodei , and Ilya Sutskever . 2019. Language models are unsupervised multitask learners. OpenAI blog, 1, 8 ( 2019 ), 9. Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019. Language models are unsupervised multitask learners. OpenAI blog, 1, 8 (2019), 9."},{"key":"e_1_3_2_1_52_1","unstructured":"Colin Raffel Noam Shazeer Adam Roberts Katherine Lee Sharan Narang Michael Matena Yanqi Zhou Wei Li and Peter J Liu. 2019. Exploring the limits of transfer learning with a unified text-to-text transformer. arXiv preprint arXiv:1910.10683. \t\t\t\t  Colin Raffel Noam Shazeer Adam Roberts Katherine Lee Sharan Narang Michael Matena Yanqi Zhou Wei Li and Peter J Liu. 2019. Exploring the limits of transfer learning with a unified text-to-text transformer. arXiv preprint arXiv:1910.10683."},{"key":"e_1_3_2_1_53_1","unstructured":"Sachin Ravi and Hugo Larochelle. 2016. Optimization as a model for few-shot learning. \t\t\t\t  Sachin Ravi and Hugo Larochelle. 2016. Optimization as a model for few-shot learning."},{"key":"e_1_3_2_1_54_1","doi-asserted-by":"publisher","DOI":"10.1145\/2884781.2884848"},{"key":"e_1_3_2_1_55_1","doi-asserted-by":"publisher","DOI":"10.1145\/2666356.2594321"},{"key":"e_1_3_2_1_56_1","unstructured":"Shuo Ren Daya Guo Shuai Lu Long Zhou Shujie Liu Duyu Tang Ming Zhou Ambrosio Blanco and Shuai Ma. 2020. CodeBLEU: a Method for Automatic Evaluation of Code Synthesis. arXiv preprint arXiv:2009.10297 arxiv:2009.10297 \t\t\t\t  Shuo Ren Daya Guo Shuai Lu Long Zhou Shujie Liu Duyu Tang Ming Zhou Ambrosio Blanco and Shuai Ma. 2020. CodeBLEU: a Method for Automatic Evaluation of Code Synthesis. arXiv preprint arXiv:2009.10297 arxiv:2009.10297"},{"key":"e_1_3_2_1_57_1","volume-title":"International conference on machine learning. 2152\u20132161","author":"Romera-Paredes Bernardino","year":"2015","unstructured":"Bernardino Romera-Paredes and Philip Torr . 2015 . An embarrassingly simple approach to zero-shot learning . In International conference on machine learning. 2152\u20132161 . Bernardino Romera-Paredes and Philip Torr. 2015. An embarrassingly simple approach to zero-shot learning. In International conference on machine learning. 2152\u20132161."},{"key":"e_1_3_2_1_58_1","volume-title":"DOBF: A deobfuscation pre-training objective for programming languages. arXiv preprint arXiv:2102.07492.","author":"Roziere Baptiste","year":"2021","unstructured":"Baptiste Roziere , Marie-Anne Lachaux , Marc Szafraniec , and Guillaume Lample . 2021 . DOBF: A deobfuscation pre-training objective for programming languages. arXiv preprint arXiv:2102.07492. Baptiste Roziere, Marie-Anne Lachaux, Marc Szafraniec, and Guillaume Lample. 2021. DOBF: A deobfuscation pre-training objective for programming languages. arXiv preprint arXiv:2102.07492."},{"key":"e_1_3_2_1_59_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00049"},{"key":"e_1_3_2_1_60_1","doi-asserted-by":"crossref","unstructured":"Michele Tufano Jevgenija Pantiuchina Cody Watson Gabriele Bavota and Denys Poshyvanyk. 2019. On Learning Meaningful Code Changes via Neural Machine Translation. arXiv preprint arXiv:1901.09102. \t\t\t\t  Michele Tufano Jevgenija Pantiuchina Cody Watson Gabriele Bavota and Denys Poshyvanyk. 2019. On Learning Meaningful Code Changes via Neural Machine Translation. arXiv preprint arXiv:1901.09102.","DOI":"10.1109\/ICSE.2019.00021"},{"key":"e_1_3_2_1_61_1","doi-asserted-by":"publisher","DOI":"10.1145\/3340544"},{"key":"e_1_3_2_1_62_1","doi-asserted-by":"publisher","DOI":"10.1145\/3106237.3106289"},{"key":"e_1_3_2_1_63_1","volume-title":"\u0141 ukasz Kaiser, and Illia Polosukhin","author":"Vaswani Ashish","year":"2017","unstructured":"Ashish Vaswani , Noam Shazeer , Niki Parmar , Jakob Uszkoreit , Llion Jones , Aidan N Gomez , \u0141 ukasz Kaiser, and Illia Polosukhin . 2017 . Attention is All you Need. In Advances in Neural Information Processing Systems 30. 5998\u20136008. Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, \u0141 ukasz Kaiser, and Illia Polosukhin. 2017. Attention is All you Need. In Advances in Neural Information Processing Systems 30. 5998\u20136008."},{"key":"e_1_3_2_1_64_1","doi-asserted-by":"crossref","unstructured":"Yue Wang Weishi Wang Shafiq Joty and Steven CH Hoi. 2021. Codet5: Identifier-aware unified pre-trained encoder-decoder models for code understanding and generation. arXiv preprint arXiv:2109.00859. \t\t\t\t  Yue Wang Weishi Wang Shafiq Joty and Steven CH Hoi. 2021. Codet5: Identifier-aware unified pre-trained encoder-decoder models for code understanding and generation. arXiv preprint arXiv:2109.00859.","DOI":"10.18653\/v1\/2021.emnlp-main.685"},{"key":"e_1_3_2_1_65_1","volume-title":"Generalizing from a few examples: A survey on few-shot learning. ACM computing surveys (csur), 53, 3","author":"Wang Yaqing","year":"2020","unstructured":"Yaqing Wang , Quanming Yao , James T Kwok , and Lionel M Ni. 2020. Generalizing from a few examples: A survey on few-shot learning. ACM computing surveys (csur), 53, 3 ( 2020 ), 1\u201334. Yaqing Wang, Quanming Yao, James T Kwok, and Lionel M Ni. 2020. Generalizing from a few examples: A survey on few-shot learning. ACM computing surveys (csur), 53, 3 (2020), 1\u201334."},{"key":"e_1_3_2_1_66_1","doi-asserted-by":"publisher","DOI":"10.1145\/2970276.2970326"},{"key":"e_1_3_2_1_67_1","volume-title":"Zero-shot learning\u2014a comprehensive evaluation of the good, the bad and the ugly","author":"Xian Yongqin","year":"2018","unstructured":"Yongqin Xian , Christoph H Lampert , Bernt Schiele , and Zeynep Akata . 2018. Zero-shot learning\u2014a comprehensive evaluation of the good, the bad and the ugly . IEEE transactions on pattern analysis and machine intelligence, 41, 9 ( 2018 ), 2251\u20132265. Yongqin Xian, Christoph H Lampert, Bernt Schiele, and Zeynep Akata. 2018. Zero-shot learning\u2014a comprehensive evaluation of the good, the bad and the ugly. IEEE transactions on pattern analysis and machine intelligence, 41, 9 (2018), 2251\u20132265."},{"key":"e_1_3_2_1_68_1","doi-asserted-by":"publisher","DOI":"10.1145\/3540250.3549094"}],"event":{"name":"ESEC\/FSE '22: 30th ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering","location":"Singapore Singapore","acronym":"ESEC\/FSE '22","sponsor":["SIGSOFT ACM Special Interest Group on Software Engineering","NUS NUS"]},"container-title":["Proceedings of the 30th ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3540250.3549162","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3540250.3549162","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3540250.3549162","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T17:51:02Z","timestamp":1750182662000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3540250.3549162"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,11,7]]},"references-count":68,"alternative-id":["10.1145\/3540250.3549162","10.1145\/3540250"],"URL":"https:\/\/doi.org\/10.1145\/3540250.3549162","relation":{},"subject":[],"published":{"date-parts":[[2022,11,7]]},"assertion":[{"value":"2022-11-09","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}