{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,21]],"date-time":"2026-04-21T15:06:31Z","timestamp":1776783991068,"version":"3.51.2"},"publisher-location":"New York, NY, USA","reference-count":42,"publisher":"ACM","license":[{"start":{"date-parts":[[2020,11,2]],"date-time":"2020-11-02T00:00:00Z","timestamp":1604275200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2020,11,2]]},"DOI":"10.1145\/3400302.3415698","type":"proceedings-article","created":{"date-parts":[[2020,12,18]],"date-time":"2020-12-18T01:17:48Z","timestamp":1608254268000},"page":"1-9","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":30,"title":["HAPI"],"prefix":"10.1145","author":[{"given":"Stefanos","family":"Laskaridis","sequence":"first","affiliation":[{"name":"Samsung AI Center, Cambridge"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Stylianos I.","family":"Venieris","sequence":"additional","affiliation":[{"name":"Samsung AI Center, Cambridge"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Hyeji","family":"Kim","sequence":"additional","affiliation":[{"name":"Samsung AI Center, Cambridge"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Nicholas D.","family":"Lane","sequence":"additional","affiliation":[{"name":"University of Cambridge"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2020,12,17]]},"reference":[{"key":"e_1_3_2_1_1_1","volume-title":"Proceedings of the 12th USENIX Conference on Operating Systems Design and Implementation (OSDI). 265--283","author":"Mart\u00edn","unstructured":"Mart\u00edn Abadi et al. 2016. TensorFlow: A System for Large-scale Machine Learning . In Proceedings of the 12th USENIX Conference on Operating Systems Design and Implementation (OSDI). 265--283 . Mart\u00edn Abadi et al. 2016. TensorFlow: A System for Large-scale Machine Learning. In Proceedings of the 12th USENIX Conference on Operating Systems Design and Implementation (OSDI). 265--283."},{"key":"e_1_3_2_1_2_1","volume-title":"EmBench: Quantifying Performance Variations of Deep Neural Networks Across Modern Commodity Devices. In International Workshop on Embedded and Mobile Deep Learning (EMDL).","author":"Almeida Mario","unstructured":"Mario Almeida , Stefanos Laskaridis , Ilias Leontiadis , Stylianos I. Venieris , and Nicholas D. Lane . 2019 . EmBench: Quantifying Performance Variations of Deep Neural Networks Across Modern Commodity Devices. In International Workshop on Embedded and Mobile Deep Learning (EMDL). Mario Almeida, Stefanos Laskaridis, Ilias Leontiadis, Stylianos I. Venieris, and Nicholas D. Lane. 2019. EmBench: Quantifying Performance Variations of Deep Neural Networks Across Modern Commodity Devices. In International Workshop on Embedded and Mobile Deep Learning (EMDL)."},{"key":"e_1_3_2_1_3_1","volume-title":"TVM: An Automated End-to-End Optimizing Compiler for Deep Learning. In 13th USENIX Symposium on Operating Systems Design and Implementation (OSDI).","author":"Tianqi","unstructured":"Tianqi Chen et al. 2018 . TVM: An Automated End-to-End Optimizing Compiler for Deep Learning. In 13th USENIX Symposium on Operating Systems Design and Implementation (OSDI). Tianqi Chen et al. 2018. TVM: An Automated End-to-End Optimizing Compiler for Deep Learning. In 13th USENIX Symposium on Operating Systems Design and Implementation (OSDI)."},{"key":"e_1_3_2_1_4_1","doi-asserted-by":"crossref","unstructured":"L. Fei-Fei J. Deng and K. Li. 2010. ImageNet: Constructing a Large-Scale Image Database. Journal of Vision (2010).  L. Fei-Fei J. Deng and K. Li. 2010. ImageNet: Constructing a Large-Scale Image Database. Journal of Vision (2010).","DOI":"10.1167\/9.8.1037"},{"key":"e_1_3_2_1_5_1","volume-title":"On Calibration of Modern Neural Networks. In International Conference on Machine Learning.","author":"Guo Chuan","unstructured":"Chuan Guo , Geoff Pleiss , Yu Sun , and Kilian Q. Weinberger . 2017 . On Calibration of Modern Neural Networks. In International Conference on Machine Learning. Chuan Guo, Geoff Pleiss, Yu Sun, and Kilian Q. Weinberger. 2017. On Calibration of Modern Neural Networks. In International Conference on Machine Learning."},{"key":"e_1_3_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1145\/2906388.2906396"},{"key":"e_1_3_2_1_7_1","unstructured":"K. He et al. 2018. Mask R-CNN. IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) (2018).  K. He et al. 2018. Mask R-CNN. IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) (2018)."},{"key":"e_1_3_2_1_8_1","volume-title":"Deep Residual Learning for Image Recognition. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR).","author":"He K.","unstructured":"K. He , X. Zhang , S. Ren , and J. Sun . 2016 . Deep Residual Learning for Image Recognition. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR). K. He, X. Zhang, S. Ren, and J. Sun. 2016. Deep Residual Learning for Image Recognition. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)."},{"key":"e_1_3_2_1_9_1","volume-title":"Focus: Querying Large Video Datasets with Low Latency and Low Cost. In 13th USENIX Conference on Operating Systems Design and Implementation (OSDI).","author":"Kevin","unstructured":"Kevin Hsieh et al. 2018 . Focus: Querying Large Video Datasets with Low Latency and Low Cost. In 13th USENIX Conference on Operating Systems Design and Implementation (OSDI). Kevin Hsieh et al. 2018. Focus: Querying Large Video Datasets with Low Latency and Low Cost. In 13th USENIX Conference on Operating Systems Design and Implementation (OSDI)."},{"key":"e_1_3_2_1_10_1","volume-title":"Multi-Scale Dense Networks for Resource Efficient Image Classification. In International Conference on Learning Representations (ICLR).","author":"Gao","unstructured":"Gao Huang et al. 2018 . Multi-Scale Dense Networks for Resource Efficient Image Classification. In International Conference on Learning Representations (ICLR). Gao Huang et al. 2018. Multi-Scale Dense Networks for Resource Efficient Image Classification. In International Conference on Learning Representations (ICLR)."},{"key":"e_1_3_2_1_11_1","volume-title":"Speed\/Accuracy Trade-Offs for Modern Convolutional Object Detectors. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR).","author":"J. Huang","year":"2017","unstructured":"J. Huang et al. 2017 . Speed\/Accuracy Trade-Offs for Modern Convolutional Object Detectors. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR). J. Huang et al. 2017. Speed\/Accuracy Trade-Offs for Modern Convolutional Object Detectors. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)."},{"key":"e_1_3_2_1_12_1","volume-title":"Shallow-Deep Networks: Understanding and Mitigating Network Overthinking. In International Conference on Machine Learning (ICML).","author":"Kaya Yigitcan","year":"2019","unstructured":"Yigitcan Kaya , Sanghyun Hong , and Tudor Dumitras . 2019 . Shallow-Deep Networks: Understanding and Mitigating Network Overthinking. In International Conference on Machine Learning (ICML). Yigitcan Kaya, Sanghyun Hong, and Tudor Dumitras. 2019. Shallow-Deep Networks: Understanding and Mitigating Network Overthinking. In International Conference on Machine Learning (ICML)."},{"key":"e_1_3_2_1_13_1","volume-title":"2018 IEEE\/RSJ International Conference on Intelligent Robots and Systems (IROS).","author":"Kouris A.","unstructured":"A. Kouris and C. Bouganis . 2018. Learning to Fly by MySelf: A Self-Supervised CNN-Based Approach for Autonomous Navigation . In 2018 IEEE\/RSJ International Conference on Intelligent Robots and Systems (IROS). A. Kouris and C. Bouganis. 2018. Learning to Fly by MySelf: A Self-Supervised CNN-Based Approach for Autonomous Navigation. In 2018 IEEE\/RSJ International Conference on Intelligent Robots and Systems (IROS)."},{"key":"e_1_3_2_1_14_1","volume-title":"Automation Test in Europe Conference Exhibition (DATE). 1656--1661","author":"Kouris A.","unstructured":"A. Kouris , S. I. Venieris , and C. Bouganis . 2020. A Throughput-Latency Co-Optimised Cascade of Convolutional Neural Network Classifiers. In 2020 Design , Automation Test in Europe Conference Exhibition (DATE). 1656--1661 . A. Kouris, S. I. Venieris, and C. Bouganis. 2020. A Throughput-Latency Co-Optimised Cascade of Convolutional Neural Network Classifiers. In 2020 Design, Automation Test in Europe Conference Exhibition (DATE). 1656--1661."},{"key":"e_1_3_2_1_15_1","volume-title":"CascadeCNN: Pushing the Performance Limits of Quantisation in Convolutional Neural Networks. In 28th International Conference on Field Programmable Logic and Applications (FPL).","author":"Kouris A.","unstructured":"A. Kouris , S. I. Venieris , and C. S. Bouganis . 2018 . CascadeCNN: Pushing the Performance Limits of Quantisation in Convolutional Neural Networks. In 28th International Conference on Field Programmable Logic and Applications (FPL). A. Kouris, S. I. Venieris, and C. S. Bouganis. 2018. CascadeCNN: Pushing the Performance Limits of Quantisation in Convolutional Neural Networks. In 28th International Conference on Field Programmable Logic and Applications (FPL)."},{"key":"e_1_3_2_1_17_1","volume-title":"The 26th Annual International Conference on Mobile Computing and Networking (MobiCom).","author":"Laskaridis Stefanos","unstructured":"Stefanos Laskaridis , Stylianos I. Venieris , Mario Almeida , Ilias Leontiadis , and Nicholas D. Lane . 2020. SPINN: Synergistic Progressive Inference of Neural Networks over Device and Cloud . In The 26th Annual International Conference on Mobile Computing and Networking (MobiCom). Stefanos Laskaridis, Stylianos I. Venieris, Mario Almeida, Ilias Leontiadis, and Nicholas D. Lane. 2020. SPINN: Synergistic Progressive Inference of Neural Networks over Device and Cloud. In The 26th Annual International Conference on Mobile Computing and Networking (MobiCom)."},{"key":"e_1_3_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/PROC.1987.13876"},{"key":"e_1_3_2_1_19_1","volume-title":"International Conference on Learning Representations (ICLR).","author":"Lee Namhoon","year":"2019","unstructured":"Namhoon Lee , Thalaiyasingam Ajanthan , and Philip Torr . 2019 . SNIP: Single-Shot Network Pruning based on Connection Sensitivity . In International Conference on Learning Representations (ICLR). Namhoon Lee, Thalaiyasingam Ajanthan, and Philip Torr. 2019. SNIP: Single-Shot Network Pruning based on Connection Sensitivity. In International Conference on Learning Representations (ICLR)."},{"key":"e_1_3_2_1_20_1","volume-title":"MobiSR: Efficient On-Device Super-Resolution Through Heterogeneous Mobile Processors. In The 25th Annual International Conference on Mobile Computing and Networking (MobiCom).","author":"Lee Royson","unstructured":"Royson Lee , Stylianos I. Venieris , Lukasz Dudziak , Sourav Bhattacharya , and Nicholas D. Lane . 2019 . MobiSR: Efficient On-Device Super-Resolution Through Heterogeneous Mobile Processors. In The 25th Annual International Conference on Mobile Computing and Networking (MobiCom). Royson Lee, Stylianos I. Venieris, Lukasz Dudziak, Sourav Bhattacharya, and Nicholas D. Lane. 2019. MobiSR: Efficient On-Device Super-Resolution Through Heterogeneous Mobile Processors. In The 25th Annual International Conference on Mobile Computing and Networking (MobiCom)."},{"key":"e_1_3_2_1_21_1","volume-title":"Improved Techniques for Training Adaptive Deep Networks. In IEEE International Conference on Computer Vision (ICCV).","author":"Li Hao","year":"2019","unstructured":"Hao Li , Hong Zhang , Xiaojuan Qi , Ruigang Yang , and Gao Huang . 2019 . Improved Techniques for Training Adaptive Deep Networks. In IEEE International Conference on Computer Vision (ICCV). Hao Li, Hong Zhang, Xiaojuan Qi, Ruigang Yang, and Gao Huang. 2019. Improved Techniques for Training Adaptive Deep Networks. In IEEE International Conference on Computer Vision (ICCV)."},{"key":"e_1_3_2_1_22_1","volume-title":"Survey of multi-objective optimization methods for engineering. Structural and multidisciplinary optimization","author":"Timothy Marler R","year":"2004","unstructured":"R Timothy Marler and Jasbir S Arora . 2004. Survey of multi-objective optimization methods for engineering. Structural and multidisciplinary optimization ( 2004 ). R Timothy Marler and Jasbir S Arora. 2004. Survey of multi-objective optimization methods for engineering. Structural and multidisciplinary optimization (2004)."},{"key":"e_1_3_2_1_23_1","volume-title":"The weighted sum method for multi-objective optimization: new insights. Structural and multidisciplinary optimization 41, 6","author":"Timothy Marler R","year":"2010","unstructured":"R Timothy Marler and Jasbir S Arora . 2010. The weighted sum method for multi-objective optimization: new insights. Structural and multidisciplinary optimization 41, 6 ( 2010 ), 853--862. R Timothy Marler and Jasbir S Arora. 2010. The weighted sum method for multi-objective optimization: new insights. Structural and multidisciplinary optimization 41, 6 (2010), 853--862."},{"key":"e_1_3_2_1_24_1","unstructured":"Colin R. Reeves (Ed.). 1993. Modern Heuristic Techniques for Combinatorial Problems. John Wiley & Sons Inc.  Colin R. Reeves (Ed.). 1993. Modern Heuristic Techniques for Combinatorial Problems. John Wiley & Sons Inc."},{"key":"e_1_3_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00474"},{"key":"e_1_3_2_1_26_1","volume-title":"Very Deep Convolutional Networks for Large-Scale Image Recognition. In International Conference on Learning Representations (ICLR).","author":"Simonyan K.","unstructured":"K. Simonyan and A. Zisserman . 2015 . Very Deep Convolutional Networks for Large-Scale Image Recognition. In International Conference on Learning Representations (ICLR). K. Simonyan and A. Zisserman. 2015. Very Deep Convolutional Networks for Large-Scale Image Recognition. In International Conference on Learning Representations (ICLR)."},{"key":"e_1_3_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1145\/3297858.3304072"},{"key":"e_1_3_2_1_28_1","volume-title":"Proc. of the IEEE","author":"Sze V.","year":"2017","unstructured":"V. Sze , Y. Chen , T. Yang , and J. S. Erner . 2017 . Efficient Processing of Deep Neural Networks: A Tutorial and Survey. Proc. of the IEEE ( 2017 ). V. Sze, Y. Chen, T. Yang, and J. S. Erner. 2017. Efficient Processing of Deep Neural Networks: A Tutorial and Survey. Proc. of the IEEE (2017)."},{"key":"e_1_3_2_1_29_1","doi-asserted-by":"crossref","unstructured":"Christian Szegedy Sergey Ioffe Vincent Vanhoucke and Alexander Alemi. 2017. Inception-v4 Inception-ResNet and the Impact of Residual Connections on Learning. In AAAI.  Christian Szegedy Sergey Ioffe Vincent Vanhoucke and Alexander Alemi. 2017. Inception-v4 Inception-ResNet and the Impact of Residual Connections on Learning. In AAAI.","DOI":"10.1609\/aaai.v31i1.11231"},{"key":"e_1_3_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1145\/3211332.3211336"},{"key":"e_1_3_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICPR.2016.7900006"},{"key":"e_1_3_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1145\/2908080.2908105"},{"key":"e_1_3_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.1109\/TNNLS.2018.2844093"},{"key":"e_1_3_2_1_34_1","volume-title":"HAQ: Hardware-Aware Automated Quantization with Mixed Precision. In CVPR.","author":"Wang Kuan","year":"2019","unstructured":"Kuan Wang , Zhijian Liu , Yujun Lin , Ji Lin , and Song Han . 2019 . HAQ: Hardware-Aware Automated Quantization with Mixed Precision. In CVPR. Kuan Wang, Zhijian Liu, Yujun Lin, Ji Lin, and Song Han. 2019. HAQ: Hardware-Aware Automated Quantization with Mixed Precision. In CVPR."},{"key":"e_1_3_2_1_35_1","doi-asserted-by":"crossref","unstructured":"S. Wang G. Ananthanarayanan Y. Zeng N. Goel A. Pathania and T. Mitra. 2019. High-Throughput CNN Inference on Embedded ARM big.LITTLE Multi-Core Processors. IEEE Transactions on Computer-Aided Design of Integrated Circuits and Systems (TCAD) (2019).  S. Wang G. Ananthanarayanan Y. Zeng N. Goel A. Pathania and T. Mitra. 2019. High-Throughput CNN Inference on Embedded ARM big.LITTLE Multi-Core Processors. IEEE Transactions on Computer-Aided Design of Integrated Circuits and Systems (TCAD) (2019).","DOI":"10.1109\/TCAD.2019.2944584"},{"key":"e_1_3_2_1_36_1","volume-title":"IEEE International Symposium on High Performance Computer Architecture (HPCA).","author":"C. Wu","year":"2019","unstructured":"C. Wu et al. 2019 . Machine Learning at Facebook: Understanding Inference at the Edge . In IEEE International Symposium on High Performance Computer Architecture (HPCA). C. Wu et al. 2019. Machine Learning at Facebook: Understanding Inference at the Edge. In IEEE International Symposium on High Performance Computer Architecture (HPCA)."},{"key":"e_1_3_2_1_37_1","volume-title":"Proceedings of the 12th USENIX Conference on Operating Systems Design and Implementation (OSDI).","author":"Wencong","unstructured":"Wencong Xiao et al. 2018. Gandiva: Introspective Cluster Scheduling for Deep Learning . In Proceedings of the 12th USENIX Conference on Operating Systems Design and Implementation (OSDI). Wencong Xiao et al. 2018. Gandiva: Introspective Cluster Scheduling for Deep Learning. In Proceedings of the 12th USENIX Conference on Operating Systems Design and Implementation (OSDI)."},{"key":"e_1_3_2_1_38_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2020.acl-main.204"},{"key":"e_1_3_2_1_39_1","doi-asserted-by":"publisher","DOI":"10.1145\/3289602.3293972"},{"key":"e_1_3_2_1_40_1","volume-title":"NetAdapt: Platform-Aware Neural Network Adaptation for Mobile Applications. In European Conference on Computer Vision (ECCV).","author":"Yang Tien-Ju","year":"2018","unstructured":"Tien-Ju Yang , Andrew Howard , Bo Chen , Xiao Zhang , Alec Go , Mark Sandler , Vivienne Sze , and Hartwig Adam . 2018 . NetAdapt: Platform-Aware Neural Network Adaptation for Mobile Applications. In European Conference on Computer Vision (ECCV). Tien-Ju Yang, Andrew Howard, Bo Chen, Xiao Zhang, Alec Go, Mark Sandler, Vivienne Sze, and Hartwig Adam. 2018. NetAdapt: Platform-Aware Neural Network Adaptation for Mobile Applications. In European Conference on Computer Vision (ECCV)."},{"key":"e_1_3_2_1_41_1","volume-title":"International Conference on Learning Representations (ICLR).","author":"Zagoruyko Sergey","year":"2017","unstructured":"Sergey Zagoruyko and Nikos Komodakis . 2017 . Paying more Attention to Attention: Improving the Performance of Convolutional Neural Networks via Attention Transfer . In International Conference on Learning Representations (ICLR). Sergey Zagoruyko and Nikos Komodakis. 2017. Paying more Attention to Attention: Improving the Performance of Convolutional Neural Networks via Attention Transfer. In International Conference on Learning Representations (ICLR)."},{"key":"e_1_3_2_1_42_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00381"},{"key":"e_1_3_2_1_43_1","volume-title":"SCAN: A Scalable Neural Networks Framework Towards Compact and Efficient Models. In Advances in Neural Information Processing Systems (NeurIPS).","author":"Zhang Linfeng","year":"2019","unstructured":"Linfeng Zhang , Zhanhong Tan , Jiebo Song , Jingwei Chen , Chenglong Bao , and Kaisheng Ma . 2019 . SCAN: A Scalable Neural Networks Framework Towards Compact and Efficient Models. In Advances in Neural Information Processing Systems (NeurIPS). Linfeng Zhang, Zhanhong Tan, Jiebo Song, Jingwei Chen, Chenglong Bao, and Kaisheng Ma. 2019. SCAN: A Scalable Neural Networks Framework Towards Compact and Efficient Models. In Advances in Neural Information Processing Systems (NeurIPS)."}],"event":{"name":"ICCAD '20: IEEE\/ACM International Conference on Computer-Aided Design","location":"Virtual Event USA","acronym":"ICCAD '20","sponsor":["SIGDA ACM Special Interest Group on Design Automation","IEEE CAS","IEEE CEDA","IEEE CS"]},"container-title":["Proceedings of the 39th International Conference on Computer-Aided Design"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3400302.3415698","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3400302.3415698","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T22:02:52Z","timestamp":1750197772000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3400302.3415698"}},"subtitle":["hardware-aware progressive inference"],"short-title":[],"issued":{"date-parts":[[2020,11,2]]},"references-count":42,"alternative-id":["10.1145\/3400302.3415698","10.1145\/3400302"],"URL":"https:\/\/doi.org\/10.1145\/3400302.3415698","relation":{},"subject":[],"published":{"date-parts":[[2020,11,2]]},"assertion":[{"value":"2020-12-17","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}