{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,9]],"date-time":"2026-07-09T15:33:26Z","timestamp":1783611206956,"version":"3.55.0"},"publisher-location":"New York, NY, USA","reference-count":17,"publisher":"ACM","license":[{"start":{"date-parts":[[2018,11,2]],"date-time":"2018-11-02T00:00:00Z","timestamp":1541116800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2018,11,2]]},"DOI":"10.1145\/3290480.3290494","type":"proceedings-article","created":{"date-parts":[[2018,12,6]],"date-time":"2018-12-06T13:37:41Z","timestamp":1544103461000},"page":"74-78","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":33,"title":["Enhancing Machine Learning Based Malware Detection Model by Reinforcement Learning"],"prefix":"10.1145","author":[{"given":"Cangshuai","family":"Wu","sequence":"first","affiliation":[{"name":"National University of Defense, Technology, ChangSha, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Jiangyong","family":"Shi","sequence":"additional","affiliation":[{"name":"National University of Defense, Technology, ChangSha, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yuexiang","family":"Yang","sequence":"additional","affiliation":[{"name":"National University of Defense, Technology, ChangSha, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Wenhua","family":"Li","sequence":"additional","affiliation":[{"name":"National University of Defense, Technology, ChangSha, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2018,11,2]]},"reference":[{"key":"e_1_3_2_1_1_1","volume-title":"Learning to evade static PE machine learning malware models via reinforcement learning{J}. arXiv preprint arXiv:1801.08917v2","author":"Anderson HS","year":"2018","unstructured":"Anderson HS , Kharkar A , Filar B , Evans D , Roth P Learning to evade static PE machine learning malware models via reinforcement learning{J}. arXiv preprint arXiv:1801.08917v2 , 2018 . Anderson HS, Kharkar A, Filar B, Evans D, Roth P et al. Learning to evade static PE machine learning malware models via reinforcement learning{J}. arXiv preprint arXiv:1801.08917v2, 2018."},{"key":"e_1_3_2_1_2_1","volume-title":"EMBER: an open dataset for training static PE maiware machine learning models{J}. arXiv preprint arXiv:1804.04637v2","author":"Anderson HS","year":"2018","unstructured":"Anderson HS , Roth P , EMBER: an open dataset for training static PE maiware machine learning models{J}. arXiv preprint arXiv:1804.04637v2 , 2018 . Anderson HS, Roth P, et al. EMBER: an open dataset for training static PE maiware machine learning models{J}. arXiv preprint arXiv:1804.04637v2, 2018."},{"key":"e_1_3_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.5555\/882495.884439"},{"key":"e_1_3_2_1_4_1","volume-title":"Adversarial deep learning for robust detection of binary encoded malware{J}. arXiv preprint arXiv:1801.02950","author":"Huang A","year":"2018","unstructured":"Huang A , Al-Dujaili A , Hemberg E , O'Reilly U M , Adversarial deep learning for robust detection of binary encoded malware{J}. arXiv preprint arXiv:1801.02950 , 2018 . Huang A, Al-Dujaili A, Hemberg E, O'Reilly U M, et al. Adversarial deep learning for robust detection of binary encoded malware{J}. arXiv preprint arXiv:1801.02950, 2018."},{"key":"e_1_3_2_1_5_1","volume-title":"January, 2009,","author":"Shafq M Z","year":"2009","unstructured":"Shafq M Z , Tabish S M , Mirza F , Farooq M , A framework for efcient mining of structural information to detect zero-day malicious portable executables. Technical report, TR-nexGINRC-2009-21 , January, 2009, available at http:\/\/www. nexginrc. org\/papers\/tr21-zubair. pdf, 2009 . Shafq M Z, Tabish S M, Mirza F, Farooq M, et al. A framework for efcient mining of structural information to detect zero-day malicious portable executables. Technical report, TR-nexGINRC-2009-21, January, 2009, available at http:\/\/www. nexginrc. org\/papers\/tr21-zubair. pdf, 2009."},{"key":"e_1_3_2_1_6_1","unstructured":"Nataraj L Karthikeyan S Jacob G Manjunath B S etal Malware images: visualization and automatic classification{J}. ISBN 987-1-4503-0679-9.  Nataraj L Karthikeyan S Jacob G Manjunath B S et al. Malware images: visualization and automatic classification{J}. ISBN 987-1-4503-0679-9."},{"key":"e_1_3_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/MALWARE.2015.7413680"},{"key":"e_1_3_2_1_8_1","volume-title":"Malware detection by eating a whole exe{J}. arXiv preprint arXiv:1710.09435","author":"Raff E","year":"2017","unstructured":"Raff E , Barker J , Sylvester J , Brandon R , Catanzaro B Nicholas C , Malware detection by eating a whole exe{J}. arXiv preprint arXiv:1710.09435 , 2017 . Raff E, Barker J, Sylvester J, Brandon R, Catanzaro B Nicholas C, et al. Malware detection by eating a whole exe{J}. arXiv preprint arXiv:1710.09435, 2017."},{"key":"e_1_3_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1145\/3195528.3195533"},{"key":"e_1_3_2_1_10_1","volume-title":"Evading classifer in the dark: Guiding unpredictable morphing using binary output blackboxes{J}. arXiv preprint arXiv:1705.07535","author":"Dang H","year":"2017","unstructured":"Dang H , Yue H , Chang E C , Evading classifer in the dark: Guiding unpredictable morphing using binary output blackboxes{J}. arXiv preprint arXiv:1705.07535 , 2017 . Dang H, Yue H, Chang E C, et al. Evading classifer in the dark: Guiding unpredictable morphing using binary output blackboxes{J}. arXiv preprint arXiv:1705.07535, 2017."},{"key":"e_1_3_2_1_11_1","volume-title":"Generating adversarial malware examples for black-box attacks based on GAN{J}. arXiv preprint arXiv:1702.05983","author":"Hu W","year":"2017","unstructured":"Hu W and Tan Y . Generating adversarial malware examples for black-box attacks based on GAN{J}. arXiv preprint arXiv:1702.05983 , 2017 . Hu W and Tan Y. Generating adversarial malware examples for black-box attacks based on GAN{J}. arXiv preprint arXiv:1702.05983, 2017."},{"key":"e_1_3_2_1_12_1","volume-title":"Reinforcement learning: An introduction","author":"Sutton R S","year":"1998","unstructured":"Sutton R S and Barto A G . Reinforcement learning: An introduction , volume 1 . MIT press Cambridge , 1998 . Sutton R S and Barto A G. Reinforcement learning: An introduction, volume 1. MIT press Cambridge, 1998."},{"key":"e_1_3_2_1_13_1","volume-title":"LIEF: library for instrumenting executable fles. https:\/\/lief.quarkslab.com\/","author":"Quarkslab","year":"2017","unstructured":"Quarkslab . LIEF: library for instrumenting executable fles. https:\/\/lief.quarkslab.com\/ , 2017 -2018. Quarkslab. LIEF: library for instrumenting executable fles. https:\/\/lief.quarkslab.com\/, 2017-2018."},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1016\/S0031-3203(96)00142-2"},{"key":"e_1_3_2_1_15_1","volume-title":"Deep reinforcement learning in large discrete action spaces{J}. arXiv preprint arXiv:1512.07679","author":"Dulac-Arnold G","year":"2015","unstructured":"Dulac-Arnold G , Evans R , Hasselt H , Sunehag P , Lillicrap T , Hunt J , Mann T , Weber T , Degris T , Coppin B Deep reinforcement learning in large discrete action spaces{J}. arXiv preprint arXiv:1512.07679 , 2015 . Dulac-Arnold G, Evans R, Hasselt H, Sunehag P, Lillicrap T, Hunt J, Mann T, Weber T, Degris T, Coppin B et al. Deep reinforcement learning in large discrete action spaces{J}. arXiv preprint arXiv:1512.07679, 2015."},{"key":"e_1_3_2_1_16_1","first-page":"3149","volume-title":"Advances in Neural Information Processing Systems","author":"Ke G","year":"2017","unstructured":"Ke G , Meng Q , Finley T , Wang T , Chen W , Ma W , Ye Q , Liu T Y , Lightgbm : A highly effcient gradient boosting decision tree{J} . In Advances in Neural Information Processing Systems , pages 3149 -- 3157 , 2017 . Ke G, Meng Q, Finley T, Wang T, Chen W, Ma W, Ye Q, Liu T Y, et al. Lightgbm: A highly effcient gradient boosting decision tree{J}. In Advances in Neural Information Processing Systems, pages 3149--3157, 2017."},{"key":"e_1_3_2_1_17_1","unstructured":"Virustotal-free online virus malware and url scanner. https:\/\/www.virustotal.com\/en. Accessed: 2018-03-09.  Virustotal-free online virus malware and url scanner. https:\/\/www.virustotal.com\/en. Accessed: 2018-03-09."}],"event":{"name":"ICCNS 2018: 2018 the 8th International Conference on Communication and Network Security","location":"Qingdao China","acronym":"ICCNS 2018"},"container-title":["Proceedings of the 8th International Conference on Communication and Network Security"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3290480.3290494","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3290480.3290494","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T00:44:17Z","timestamp":1750207457000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3290480.3290494"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2018,11,2]]},"references-count":17,"alternative-id":["10.1145\/3290480.3290494","10.1145\/3290480"],"URL":"https:\/\/doi.org\/10.1145\/3290480.3290494","relation":{},"subject":[],"published":{"date-parts":[[2018,11,2]]},"assertion":[{"value":"2018-11-02","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}