{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,28]],"date-time":"2026-06-28T04:38:18Z","timestamp":1782621498624,"version":"3.54.5"},"publisher-location":"New York, NY, USA","reference-count":65,"publisher":"ACM","license":[{"start":{"date-parts":[[2017,10,14]],"date-time":"2017-10-14T00:00:00Z","timestamp":1507939200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/100000001","name":"National Science Foundation","doi-asserted-by":"publisher","award":["CCF-XPS-1438996, CCF-XPS-1628991, and IIS-VEC-1539011."],"award-info":[{"award-number":["CCF-XPS-1438996, CCF-XPS-1628991, and IIS-VEC-1539011."]}],"id":[{"id":"10.13039\/100000001","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2017,10,14]]},"DOI":"10.1145\/3123939.3123970","type":"proceedings-article","created":{"date-parts":[[2017,11,20]],"date-time":"2017-11-20T14:31:12Z","timestamp":1511188272000},"page":"786-799","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":46,"title":["DeftNN"],"prefix":"10.1145","author":[{"given":"Parker","family":"Hill","sequence":"first","affiliation":[{"name":"University of Michigan"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Animesh","family":"Jain","sequence":"additional","affiliation":[{"name":"University of Michigan"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Mason","family":"Hill","sequence":"additional","affiliation":[{"name":"University of Michigan and University of Nevada"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Babak","family":"Zamirai","sequence":"additional","affiliation":[{"name":"University of Michigan"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Chang-Hong","family":"Hsu","sequence":"additional","affiliation":[{"name":"University of Michigan"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Michael A.","family":"Laurenzano","sequence":"additional","affiliation":[{"name":"University of Michigan"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Scott","family":"Mahlke","sequence":"additional","affiliation":[{"name":"University of Michigan"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Lingjia","family":"Tang","sequence":"additional","affiliation":[{"name":"University of Michigan"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Jason","family":"Mars","sequence":"additional","affiliation":[{"name":"University of Michigan"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2017,10,14]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISCA.2016.11"},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1145\/2744769.2744788"},{"key":"e_1_3_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1145\/1815961.1815993"},{"key":"e_1_3_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1145\/2541940.2541967"},{"key":"e_1_3_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/MICRO.2014.58"},{"key":"e_1_3_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISCA.2016.40"},{"key":"e_1_3_2_1_7_1","unstructured":"Sharan Chetlur Cliff Woolley Philippe Vandermersch Jonathan Cohen John Tran Bryan Catanzaro and Evan Shelhamer. 2014. cuDNN: Efficient primitives for deep learning. arXiv:1410.0759.  Sharan Chetlur Cliff Woolley Philippe Vandermersch Jonathan Cohen John Tran Bryan Catanzaro and Evan Shelhamer. 2014. cuDNN: Efficient primitives for deep learning. arXiv:1410.0759."},{"key":"e_1_3_2_1_8_1","volume-title":"Retrieved","author":"Collobert Ronan","year":"2015","unstructured":"Ronan Collobert , Clement Farabet , Koray Kavukcuoglu , and Soumith Chintala . 2015 . torch. (2015) . Retrieved August 25, 2017 from http:\/\/torch.ch\/ Ronan Collobert, Clement Farabet, Koray Kavukcuoglu, and Soumith Chintala. 2015. torch. (2015). Retrieved August 25, 2017 from http:\/\/torch.ch\/"},{"key":"e_1_3_2_1_9_1","volume-title":"Automation & Test in Europe Conference & Exhibition (DATE).","author":"Conti Francesco","year":"2015","unstructured":"Francesco Conti and Luca Benini . 2015 . A ultra-low-energy convolution engine for fast brain-inspired vision in multicore clusters. In Design , Automation & Test in Europe Conference & Exhibition (DATE). Francesco Conti and Luca Benini. 2015. A ultra-low-energy convolution engine for fast brain-inspired vision in multicore clusters. In Design, Automation & Test in Europe Conference & Exhibition (DATE)."},{"key":"e_1_3_2_1_10_1","unstructured":"Matthieu Courbariaux Yoshua Bengio and Jean-Pierre David. 2014. Low precision arithmetic for deep learning. arXiv:1412.7024.  Matthieu Courbariaux Yoshua Bengio and Jean-Pierre David. 2014. Low precision arithmetic for deep learning. arXiv:1412.7024."},{"key":"e_1_3_2_1_11_1","volume-title":"Andrew Senior, Paul Tucker, Ke Yang, Quoc V. Le, and Andrew Y. Ng.","author":"Dean Jeffrey","year":"2012","unstructured":"Jeffrey Dean , Greg Corrado , Rajat Monga , Kai Chen , Matthieu Devin , Mark Mao , Marc extquotesingle aurelio Ranzato , Andrew Senior, Paul Tucker, Ke Yang, Quoc V. Le, and Andrew Y. Ng. 2012 . Large scale distributed deep networks. In Neural Information Processing Systems (NIPS) . Jeffrey Dean, Greg Corrado, Rajat Monga, Kai Chen, Matthieu Devin, Mark Mao, Marc extquotesingle aurelio Ranzato, Andrew Senior, Paul Tucker, Ke Yang, Quoc V. Le, and Andrew Y. Ng. 2012. Large scale distributed deep networks. In Neural Information Processing Systems (NIPS)."},{"key":"e_1_3_2_1_12_1","volume-title":"Asia and South Pacific Design Automation Conference (ASP-DAC).","author":"Du Zidong","year":"2014","unstructured":"Zidong Du , Avinash Lingamneni , Yunji Chen , Krishna Palem , Olivier Temam , and Chengyong Wu . 2014 . Leveraging the error resilience of machine-learning applications for designing highly energy efficient accelerators . In Asia and South Pacific Design Automation Conference (ASP-DAC). Zidong Du, Avinash Lingamneni, Yunji Chen, Krishna Palem, Olivier Temam, and Chengyong Wu. 2014. Leveraging the error resilience of machine-learning applications for designing highly energy efficient accelerators. In Asia and South Pacific Design Automation Conference (ASP-DAC)."},{"key":"e_1_3_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPRW.2011.5981829"},{"key":"e_1_3_2_1_14_1","volume-title":"Retrieved","author":"Finley Klint","year":"2015","unstructured":"Klint Finley . 2015 . Facebook open-sources a trove of AI tools. (2015) . Retrieved August 25, 2017 from https:\/\/www.wired.com\/2015\/01\/facebook-open-sources-trove-ai-tools\/ Klint Finley. 2015. Facebook open-sources a trove of AI tools. (2015). Retrieved August 25, 2017 from https:\/\/www.wired.com\/2015\/01\/facebook-open-sources-trove-ai-tools\/"},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1109\/MICRO.2007.12"},{"key":"e_1_3_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2015.169"},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1145\/2694344.2694351"},{"key":"e_1_3_2_1_18_1","volume-title":"Retrieved","year":"2015","unstructured":"Google. 2015 . TensorFlow. (2015) . Retrieved August 25, 2017 from http:\/\/www.tensorflow.org\/ Google. 2015. TensorFlow. (2015). Retrieved August 25, 2017 from http:\/\/www.tensorflow.org\/"},{"key":"e_1_3_2_1_19_1","unstructured":"Suyog Gupta Ankur Agrawal Kailash Gopalakrishnan and Pritish Narayanan. 2015. Deep Learning with Limited Numerical Precision. arXiv:1502.02551.  Suyog Gupta Ankur Agrawal Kailash Gopalakrishnan and Pritish Narayanan. 2015. Deep Learning with Limited Numerical Precision. arXiv:1502.02551."},{"key":"e_1_3_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISCA.2016.30"},{"key":"e_1_3_2_1_21_1","volume-title":"International Conference on Learning Representations (ICLR).","author":"Han Song","year":"2015","unstructured":"Song Han , Huizi Mao , and William J Dally . 2015 . Deep compression: Compressing deep neural network with pruning, trained quantization and huffman coding . International Conference on Learning Representations (ICLR). Song Han, Huizi Mao, and William J Dally. 2015. Deep compression: Compressing deep neural network with pruning, trained quantization and huffman coding. International Conference on Learning Representations (ICLR)."},{"key":"e_1_3_2_1_22_1","unstructured":"Song Han Jeff Pool John Tran and William Dally. 2015. Learning both Weights and Connections for Efficient Neural Network. Neural Information Processing Systems (NIPS).   Song Han Jeff Pool John Tran and William Dally. 2015. Learning both Weights and Connections for Efficient Neural Network. Neural Information Processing Systems (NIPS)."},{"key":"e_1_3_2_1_23_1","volume-title":"Ng","author":"Hannun Awni Y.","year":"2014","unstructured":"Awni Y. Hannun , Carl Case , Jared Casper , Bryan Catanzaro , Greg Diamos , Erich Elsen , Ryan Prenger , Sanjeev Satheesh , Shubho Sengupta , Adam Coates , and Andrew Y . Ng . 2014 . DeepSpeech : Scaling up end-to-end speech recognition. arXiv:1412.5567. Awni Y. Hannun, Carl Case, Jared Casper, Bryan Catanzaro, Greg Diamos, Erich Elsen, Ryan Prenger, Sanjeev Satheesh, Shubho Sengupta, Adam Coates, and Andrew Y. Ng. 2014. DeepSpeech: Scaling up end-to-end speech recognition. arXiv:1412.5567."},{"key":"e_1_3_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1145\/2749469.2749472"},{"key":"e_1_3_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1145\/2694344.2694347"},{"key":"e_1_3_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1145\/1555754.1555775"},{"key":"e_1_3_2_1_27_1","volume-title":"Retrieved","year":"2015","unstructured":"Intel. 2015 . neon. (2015) . Retrieved August 25, 2017 from https:\/\/github.com\/NervanaSystems\/neon Intel. 2015. neon. (2015). Retrieved August 25, 2017 from https:\/\/github.com\/NervanaSystems\/neon"},{"key":"e_1_3_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1109\/MICRO.2016.7783744"},{"key":"e_1_3_2_1_29_1","volume-title":"Caffe: Convolutional Architecture for Fast Feature Embedding. arXiv:1408.5093.","author":"Jia Yangqing","year":"2014","unstructured":"Yangqing Jia , Evan Shelhamer , Jeff Donahue , Sergey Karayev , Jonathan Long , Ross Girshick , Sergio Guadarrama , and Trevor Darrell . 2014 . Caffe: Convolutional Architecture for Fast Feature Embedding. arXiv:1408.5093. Yangqing Jia, Evan Shelhamer, Jeff Donahue, Sergey Karayev, Jonathan Long, Ross Girshick, Sergio Guadarrama, and Trevor Darrell. 2014. Caffe: Convolutional Architecture for Fast Feature Embedding. arXiv:1408.5093."},{"key":"e_1_3_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1145\/2925426.2926294"},{"key":"e_1_3_2_1_31_1","doi-asserted-by":"crossref","unstructured":"Sergey Karayev Matthew Trentacoste Helen Han Aseem Agarwala Trevor Darrell Aaron Hertzmann and Holger Winnemoeller. 2013. Recognizing image style. arXiv:1311.3715.  Sergey Karayev Matthew Trentacoste Helen Han Aseem Agarwala Trevor Darrell Aaron Hertzmann and Holger Winnemoeller. 2013. Recognizing image style. arXiv:1311.3715.","DOI":"10.5244\/C.28.122"},{"key":"e_1_3_2_1_32_1","doi-asserted-by":"crossref","unstructured":"Joo-Young Kim Minsu Kim Seungjin Lee Jinwook Oh Kwanho Kim and Hoi-Jun Yoo. 2010. A 201.4 GOPS 496 mW real-time multi-object recognition processor with bio-inspired neural perception engine. In Journal of Solid-State Circuits (JSSC).  Joo-Young Kim Minsu Kim Seungjin Lee Jinwook Oh Kwanho Kim and Hoi-Jun Yoo. 2010. A 201.4 GOPS 496 mW real-time multi-object recognition processor with bio-inspired neural perception engine. In Journal of Solid-State Circuits (JSSC).","DOI":"10.1109\/ISSCC.2009.4977352"},{"key":"e_1_3_2_1_33_1","volume-title":"Learning multiple layers of features from tiny images. Tech report","author":"Krizhevsky Alex","unstructured":"Alex Krizhevsky and Geoffrey Hinton . 2009. Learning multiple layers of features from tiny images. Tech report , University of Toronto . Alex Krizhevsky and Geoffrey Hinton. 2009. Learning multiple layers of features from tiny images. Tech report, University of Toronto."},{"key":"e_1_3_2_1_34_1","unstructured":"Alex Krizhevsky Ilya Sutskever and Geoffrey E Hinton. 2012. Imagenet classification with deep convolutional neural networks. In Neural Information Processing Systems (NIPS).   Alex Krizhevsky Ilya Sutskever and Geoffrey E Hinton. 2012. Imagenet classification with deep convolutional neural networks. In Neural Information Processing Systems (NIPS)."},{"key":"e_1_3_2_1_35_1","volume-title":"Retrieved","author":"LISA","year":"2015","unstructured":"LISA lab. 2015 . theano. (2015) . Retrieved August 25, 2017 from http:\/\/deeplearning.net\/software\/theano\/ LISA lab. 2015. theano. (2015). Retrieved August 25, 2017 from http:\/\/deeplearning.net\/software\/theano\/"},{"key":"e_1_3_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1145\/1273496.1273556"},{"key":"e_1_3_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.1145\/2908080.2908087"},{"key":"e_1_3_2_1_38_1","unstructured":"Andrew Lavin. 2015. maxDNN: An Efficient Convolution Kernel for Deep Learning with Maxwell GPUs. arXiv:1501.06633.  Andrew Lavin. 2015. maxDNN: An Efficient Convolution Kernel for Deep Learning with Maxwell GPUs. arXiv:1501.06633."},{"key":"e_1_3_2_1_39_1","doi-asserted-by":"crossref","unstructured":"Yann LeCun Yoshua Bengio and Geoffrey Hinton. 2015. Deep learning. Nature.  Yann LeCun Yoshua Bengio and Geoffrey Hinton. 2015. Deep learning. Nature.","DOI":"10.1038\/nature14539"},{"key":"e_1_3_2_1_40_1","doi-asserted-by":"publisher","DOI":"10.1109\/5.726791"},{"key":"e_1_3_2_1_41_1","volume-title":"Retrieved","author":"LeCun Yann","year":"1998","unstructured":"Yann LeCun , Corinna Cortes , and Christopher JC Burges . 1998 . The MNIST database of handwritten digits. (1998) . Retrieved August 25, 2017 from http:\/\/yann.lecun.com\/exdb\/mnist\/ Yann LeCun, Corinna Cortes, and Christopher JC Burges. 1998. The MNIST database of handwritten digits. (1998). Retrieved August 25, 2017 from http:\/\/yann.lecun.com\/exdb\/mnist\/"},{"key":"e_1_3_2_1_42_1","doi-asserted-by":"publisher","DOI":"10.1155\/2009\/258921"},{"key":"e_1_3_2_1_43_1","doi-asserted-by":"publisher","DOI":"10.1109\/ASPDAC.2014.6742916"},{"key":"e_1_3_2_1_44_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISCA.2016.16"},{"key":"e_1_3_2_1_45_1","doi-asserted-by":"publisher","DOI":"10.1145\/2155620.2155656"},{"key":"e_1_3_2_1_46_1","doi-asserted-by":"crossref","unstructured":"Rajib Nath Stanimire Tomov and Jack Dongarra. 2010. Accelerating GPU kernels for dense linear algebra. In High Performance Computing for Computational Science (VECPAR).   Rajib Nath Stanimire Tomov and Jack Dongarra. 2010. Accelerating GPU kernels for dense linear algebra. In High Performance Computing for Computational Science (VECPAR).","DOI":"10.1007\/978-3-642-19328-6_10"},{"key":"e_1_3_2_1_47_1","volume-title":"International Conference on Machine Learning (ICML).","author":"Ngiam Jiquan","year":"2011","unstructured":"Jiquan Ngiam , Adam Coates , Ahbik Lahiri , Bobby Prochnow , Quoc V Le , and Andrew Y Ng . 2011 . On optimization methods for deep learning . In International Conference on Machine Learning (ICML). Jiquan Ngiam, Adam Coates, Ahbik Lahiri, Bobby Prochnow, Quoc V Le, and Andrew Y Ng. 2011. On optimization methods for deep learning. In International Conference on Machine Learning (ICML)."},{"key":"e_1_3_2_1_48_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICVGIP.2008.47"},{"key":"e_1_3_2_1_49_1","volume-title":"Retrieved","year":"2017","unstructured":"Nvidia. 2017 . cuBLAS. (2017) . Retrieved August 25, 2017 from developer.nvidia.com\/cublas Nvidia. 2017. cuBLAS. (2017). Retrieved August 25, 2017 from developer.nvidia.com\/cublas"},{"key":"e_1_3_2_1_50_1","volume-title":"Retrieved","year":"2017","unstructured":"Nvidia. 2017 . cuSPARSE. (2017) . Retrieved August 25, 2017 from developer.nvidia.com\/cusparse Nvidia. 2017. cuSPARSE. (2017). Retrieved August 25, 2017 from developer.nvidia.com\/cusparse"},{"key":"e_1_3_2_1_51_1","volume-title":"Retrieved","year":"2017","unstructured":"Nvidia. 2017 . GeForce GTX TITAN X, Specifications. (2017) . Retrieved August 25, 2017 from http:\/\/www.geforce.com\/hardware\/desktop-gpus\/geforce-gtx-titan-x\/specifications Nvidia. 2017. GeForce GTX TITAN X, Specifications. (2017). Retrieved August 25, 2017 from http:\/\/www.geforce.com\/hardware\/desktop-gpus\/geforce-gtx-titan-x\/specifications"},{"key":"e_1_3_2_1_52_1","volume-title":"Retrieved","year":"2017","unstructured":"Nvidia. 2017 . Parallel Thread Execution ISA Version 5.0. (2017) . Retrieved August 25, 2017 from http:\/\/docs.nvidia.com\/cuda\/parallel-thread-execution Nvidia. 2017. Parallel Thread Execution ISA Version 5.0. (2017). Retrieved August 25, 2017 from http:\/\/docs.nvidia.com\/cuda\/parallel-thread-execution"},{"key":"e_1_3_2_1_53_1","volume-title":"Retrieved","author":"Ovtcharov Kalin","year":"2015","unstructured":"Kalin Ovtcharov , Olatunji Ruwase , Joo-Young Kim , Jeremy Fowers , Karin Strauss , and Eric Chung . 2015 . Accelerating Deep Convolutional Neural Networks Using Specialized Hardware. (2015) . Retrieved August 25, 2017 from https:\/\/www.microsoft.com\/en-us\/research\/publication\/accelerating-deep-convolutional-neural-networks-using-specialized-hardware Kalin Ovtcharov, Olatunji Ruwase, Joo-Young Kim, Jeremy Fowers, Karin Strauss, and Eric Chung. 2015. Accelerating Deep Convolutional Neural Networks Using Specialized Hardware. (2015). Retrieved August 25, 2017 from https:\/\/www.microsoft.com\/en-us\/research\/publication\/accelerating-deep-convolutional-neural-networks-using-specialized-hardware"},{"key":"e_1_3_2_1_54_1","doi-asserted-by":"publisher","DOI":"10.1109\/CICC.2012.6330636"},{"key":"e_1_3_2_1_55_1","doi-asserted-by":"publisher","DOI":"10.5555\/2388996.2389070"},{"key":"e_1_3_2_1_56_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISCA.2016.32"},{"key":"e_1_3_2_1_57_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-015-0816-y"},{"key":"e_1_3_2_1_58_1","doi-asserted-by":"publisher","DOI":"10.1145\/2540708.2540711"},{"key":"e_1_3_2_1_59_1","volume-title":"Retrieved","author":"Sato Kaz","year":"2017","unstructured":"Kaz Sato , Cliff Young , and David Patterson . 2017 . An in-depth look at Google's first Tensor Processing Unit (TPU). (2017) . Retrieved August 25, 2017 from https:\/\/cloud.google.com\/blog\/big-data\/2017\/05\/an-in-depth-look-at-googles-first-tensor-processing-unit-tpu Kaz Sato, Cliff Young, and David Patterson. 2017. An in-depth look at Google's first Tensor Processing Unit (TPU). (2017). Retrieved August 25, 2017 from https:\/\/cloud.google.com\/blog\/big-data\/2017\/05\/an-in-depth-look-at-googles-first-tensor-processing-unit-tpu"},{"key":"e_1_3_2_1_60_1","doi-asserted-by":"publisher","DOI":"10.1109\/IJCNN.2008.4633828"},{"key":"e_1_3_2_1_61_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISCA.2016.12"},{"key":"e_1_3_2_1_62_1","doi-asserted-by":"publisher","DOI":"10.5555\/2337159.2337200"},{"key":"e_1_3_2_1_63_1","volume-title":"Deep Learning and Unsupervised Feature Learning Workshop.","author":"Vanhoucke Vincent","year":"2011","unstructured":"Vincent Vanhoucke , Andrew Senior , and Mark Z Mao . 2011 . Improving the speed of neural networks on CPUs . In Deep Learning and Unsupervised Feature Learning Workshop. Vincent Vanhoucke, Andrew Senior, and Mark Z Mao. 2011. Improving the speed of neural networks on CPUs. In Deep Learning and Unsupervised Feature Learning Workshop."},{"key":"e_1_3_2_1_64_1","unstructured":"Jason Yosinski Jeff Clune Yoshua Bengio and Hod Lipson. 2014. How transferable are features in deep neural networks?. In Neural Information Processing Systems (NIPS).   Jason Yosinski Jeff Clune Yoshua Bengio and Hod Lipson. 2014. How transferable are features in deep neural networks?. In Neural Information Processing Systems (NIPS)."},{"key":"e_1_3_2_1_65_1","doi-asserted-by":"crossref","unstructured":"Jianming Zhang Shuga Ma Mehrnoosh Sameki Stan Sclaroff Margrit Betke Zhe Lin Xiaohui Shen Brian Price and Radom\u00edr M\u011bch. 2015. Salient Object Subitizing. In Computer Vision and Pattern Recognition (CVPR).  Jianming Zhang Shuga Ma Mehrnoosh Sameki Stan Sclaroff Margrit Betke Zhe Lin Xiaohui Shen Brian Price and Radom\u00edr M\u011bch. 2015. Salient Object Subitizing. In Computer Vision and Pattern Recognition (CVPR).","DOI":"10.1109\/CVPR.2015.7299031"}],"event":{"name":"MICRO-50: The 50th Annual IEEE\/ACM International Symposium on Microarchitecture","location":"Cambridge Massachusetts","acronym":"MICRO-50","sponsor":["SIGMICRO ACM Special Interest Group on Microarchitectural Research and Processing","IEEE-CS\\DATC IEEE Computer Society"]},"container-title":["Proceedings of the 50th Annual IEEE\/ACM International Symposium on Microarchitecture"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3123939.3123970","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3123939.3123970","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3123939.3123970","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T03:30:31Z","timestamp":1750217431000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3123939.3123970"}},"subtitle":["addressing bottlenecks for DNN execution on GPUs via synapse vector elimination and near-compute data fission"],"short-title":[],"issued":{"date-parts":[[2017,10,14]]},"references-count":65,"alternative-id":["10.1145\/3123939.3123970","10.1145\/3123939"],"URL":"https:\/\/doi.org\/10.1145\/3123939.3123970","relation":{},"subject":[],"published":{"date-parts":[[2017,10,14]]},"assertion":[{"value":"2017-10-14","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}