{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T23:40:09Z","timestamp":1750203609391,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":18,"publisher":"ACM","license":[{"start":{"date-parts":[[2020,4,20]],"date-time":"2020-04-20T00:00:00Z","timestamp":1587340800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2020,4,20]]},"DOI":"10.1145\/3358960.3379143","type":"proceedings-article","created":{"date-parts":[[2020,5,4]],"date-time":"2020-05-04T07:48:26Z","timestamp":1588578506000},"page":"202-209","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["DLBricks: Composable Benchmark Generation to Reduce Deep Learning Benchmarking Effort on CPUs"],"prefix":"10.1145","author":[{"given":"Cheng","family":"Li","sequence":"first","affiliation":[{"name":"University of Illinois Urbana-Champaign, Urbana, IL, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Abdul","family":"Dakkak","sequence":"additional","affiliation":[{"name":"University of Illinois Urbana-Champaign, URBANA, IL, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jinjun","family":"Xiong","sequence":"additional","affiliation":[{"name":"IBM T. J. Watson Research Center, Yorktown Heights, NY, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Wen-mei","family":"Hwu","sequence":"additional","affiliation":[{"name":"University of Illinois Urbana-Champaign, Urbana, IL, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2020,4,20]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/IISWC.2016.7581275"},{"key":"e_1_3_2_1_2_1","unstructured":"Amazon. 2019. Recommended CPU Instances. docs.aws.com\/dlami\/latest\/devguide\/cpu.html. Accessed: 2019--10--17.  Amazon. 2019. Recommended CPU Instances. docs.aws.com\/dlami\/latest\/devguide\/cpu.html. Accessed: 2019--10--17."},{"key":"e_1_3_2_1_3_1","unstructured":"Baidu. 2019. DeepBench. github.com\/baidu-research\/DeepBench.  Baidu. 2019. DeepBench. github.com\/baidu-research\/DeepBench."},{"key":"e_1_3_2_1_4_1","unstructured":"Soumith Chintala. 2019. ConvNet Benchmarks. github.com\/soumith\/convnet-benchmarks .  Soumith Chintala. 2019. ConvNet Benchmarks. github.com\/soumith\/convnet-benchmarks ."},{"key":"e_1_3_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/MM.2018.112130030"},{"key":"e_1_3_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1145\/3184407.3184423"},{"key":"e_1_3_2_1_7_1","first-page":"1","article-title":"Neural Architecture Search: A Survey","volume":"20","author":"Elsken Thomas","year":"2019","unstructured":"Thomas Elsken , Jan Hendrik Metzen , and Frank Hutter . 2019 . Neural Architecture Search: A Survey . Journal of Machine Learning Research , Vol. 20 , 55 (2019), 1 -- 21 . Thomas Elsken, Jan Hendrik Metzen, and Frank Hutter. 2019. Neural Architecture Search: A Survey. Journal of Machine Learning Research, Vol. 20, 55 (2019), 1--21.","journal-title":"Journal of Machine Learning Research"},{"volume-title":"2018 IEEE International Symposium on High Performance Computer Architecture (HPCA). IEEE, IEEE, 620--629","author":"Kim","key":"e_1_3_2_1_8_1","unstructured":"Kim Hazelwood and et al. 2018. Applied Machine Learning at Facebook: A Datacenter Infrastructure Perspective . In 2018 IEEE International Symposium on High Performance Computer Architecture (HPCA). IEEE, IEEE, 620--629 . Kim Hazelwood and et al. 2018. Applied Machine Learning at Facebook: A Datacenter Infrastructure Perspective. In 2018 IEEE International Symposium on High Performance Computer Architecture (HPCA). IEEE, IEEE, 620--629."},{"key":"e_1_3_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1109\/TCAD.2002.800456"},{"key":"e_1_3_2_1_10_1","volume-title":"Latency and Inform Optimizations of Deep Learning Models on GPUs. IEEE. The 34th IEEE International Parallel & Distributed Processing Symposium (IPDPS'20)","author":"Li Cheng","year":"2020","unstructured":"Cheng Li , Abdul Dakkak , Jinjun Xiong , and Wen-Mei Hwu . 2020 . Benanza: Automatic \u03bcBenchmark Generation to Compute \u201cLower-bound \u201d Latency and Inform Optimizations of Deep Learning Models on GPUs. IEEE. The 34th IEEE International Parallel & Distributed Processing Symposium (IPDPS'20) . Cheng Li, Abdul Dakkak, Jinjun Xiong, and Wen-Mei Hwu. 2020. Benanza: Automatic \u03bcBenchmark Generation to Compute \u201cLower-bound\u201d Latency and Inform Optimizations of Deep Learning Models on GPUs. IEEE. The 34th IEEE International Parallel & Distributed Processing Symposium (IPDPS'20)."},{"key":"e_1_3_2_1_11_1","unstructured":"MLPerf. 2019. MLPerf. github.com\/mlperf .  MLPerf. 2019. MLPerf. github.com\/mlperf ."},{"key":"e_1_3_2_1_12_1","unstructured":"Scopus Preview. [n.d.]. Scopus Preview. https:\/\/www.scopus.com\/. Accessed: 2019--10--17.  Scopus Preview. [n.d.]. Scopus Preview. https:\/\/www.scopus.com\/. Accessed: 2019--10--17."},{"key":"e_1_3_2_1_13_1","volume-title":"Very Deep Convolutional Networks for Large-Scale Image Recognition. CoRR","author":"Simonyan Karen","year":"2014","unstructured":"Karen Simonyan and Andrew Zisserman . 2014. Very Deep Convolutional Networks for Large-Scale Image Recognition. CoRR , Vol. abs\/ 1409 .1556 ( 2014 ). arxiv.org\/abs\/1409.1556 Karen Simonyan and Andrew Zisserman. 2014. Very Deep Convolutional Networks for Large-Scale Image Recognition. CoRR, Vol. abs\/1409.1556 (2014). arxiv.org\/abs\/1409.1556"},{"key":"e_1_3_2_1_14_1","volume-title":"Rethinking the Inception Architecture for Computer Vision. In 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR). IEEE, 2818--2826","author":"Szegedy Christian","year":"2016","unstructured":"Christian Szegedy , Vincent Vanhoucke , Sergey Ioffe , Jon Shlens , and Zbigniew Wojna . 2016 . Rethinking the Inception Architecture for Computer Vision. In 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR). IEEE, 2818--2826 . Christian Szegedy, Vincent Vanhoucke, Sergey Ioffe, Jon Shlens, and Zbigniew Wojna. 2016. Rethinking the Inception Architecture for Computer Vision. In 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR). IEEE, 2818--2826."},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1186\/s40537-016-0043-6"},{"key":"e_1_3_2_1_16_1","volume-title":"FBNet: Hardware-aware Efficient ConvNet Design via Differentiable Neural Architecture Search. CoRR","author":"Wu Bichen","year":"2018","unstructured":"Bichen Wu , Xiaoliang Dai , Peizhao Zhang , Yanghan Wang , Fei Sun , Yiming Wu , Yuandong Tian , Peter Vajda , Yangqing Jia , and Kurt Keutzer . 2018. FBNet: Hardware-aware Efficient ConvNet Design via Differentiable Neural Architecture Search. CoRR , Vol. abs\/ 1812 .03443 ( 2018 ). arxiv.org\/abs\/1812.03443 Bichen Wu, Xiaoliang Dai, Peizhao Zhang, Yanghan Wang, Fei Sun, Yiming Wu, Yuandong Tian, Peter Vajda, Yangqing Jia, and Kurt Keutzer. 2018. FBNet: Hardware-aware Efficient ConvNet Design via Differentiable Neural Architecture Search. CoRR, Vol. abs\/1812.03443 (2018). arxiv.org\/abs\/1812.03443"},{"key":"e_1_3_2_1_17_1","volume-title":"AI Matrix: A Deep Learning Benchmark for Alibaba Data Centers. arXiv preprint arXiv:1909.10562","author":"Zhang Wei","year":"2019","unstructured":"Wei Zhang , Wei Wei , Lingjie Xu , Lingling Jin , and Cheng Li. 2019. AI Matrix: A Deep Learning Benchmark for Alibaba Data Centers. arXiv preprint arXiv:1909.10562 ( 2019 ). Wei Zhang, Wei Wei, Lingjie Xu, Lingling Jin, and Cheng Li. 2019. AI Matrix: A Deep Learning Benchmark for Alibaba Data Centers. arXiv preprint arXiv:1909.10562 (2019)."},{"key":"e_1_3_2_1_18_1","volume-title":"Benchmarking and Analyzing Deep Neural Network Training. In 2018 IEEE International Symposium on Workload Characterization (IISWC). IEEE, IEEE, 88--100","author":"Zhu Hongyu","year":"2018","unstructured":"Hongyu Zhu , Mohamed Akrout , Bojian Zheng , Andrew Pelegris , Anand Jayarajan , Amar Phanishayee , Bianca Schroeder , and Gennady Pekhimenko . 2018 . Benchmarking and Analyzing Deep Neural Network Training. In 2018 IEEE International Symposium on Workload Characterization (IISWC). IEEE, IEEE, 88--100 . Hongyu Zhu, Mohamed Akrout, Bojian Zheng, Andrew Pelegris, Anand Jayarajan, Amar Phanishayee, Bianca Schroeder, and Gennady Pekhimenko. 2018. Benchmarking and Analyzing Deep Neural Network Training. In 2018 IEEE International Symposium on Workload Characterization (IISWC). IEEE, IEEE, 88--100."}],"event":{"name":"ICPE '20: ACM\/SPEC International Conference on Performance Engineering","sponsor":["SIGMETRICS ACM Special Interest Group on Measurement and Evaluation","SIGSOFT ACM Special Interest Group on Software Engineering"],"location":"Edmonton AB Canada","acronym":"ICPE '20"},"container-title":["Proceedings of the ACM\/SPEC International Conference on Performance Engineering"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3358960.3379143","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3358960.3379143","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T23:23:39Z","timestamp":1750202619000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3358960.3379143"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,4,20]]},"references-count":18,"alternative-id":["10.1145\/3358960.3379143","10.1145\/3358960"],"URL":"https:\/\/doi.org\/10.1145\/3358960.3379143","relation":{},"subject":[],"published":{"date-parts":[[2020,4,20]]},"assertion":[{"value":"2020-04-20","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}