{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,9]],"date-time":"2026-06-09T16:12:24Z","timestamp":1781021544052,"version":"3.54.1"},"publisher-location":"New York, NY, USA","reference-count":63,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,10,12]],"date-time":"2022-10-12T00:00:00Z","timestamp":1665532800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,10,12]]},"DOI":"10.1145\/3564121.3564796","type":"proceedings-article","created":{"date-parts":[[2023,5,16]],"date-time":"2023-05-16T19:59:41Z","timestamp":1684267181000},"page":"1-8","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":5,"title":["Automated Deep Learning Model Partitioning for Heterogeneous Edge Devices"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0000-0001-5052-4476","authenticated-orcid":false,"given":"Arijit","family":"Mukherjee","sequence":"first","affiliation":[{"name":"Embedded Devices &amp; Intelligent Systems, TCS Research, IN"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-3988-1445","authenticated-orcid":false,"given":"Swarnava","family":"Dey","sequence":"additional","affiliation":[{"name":"Embedded Devices &amp; Intelligent Systems, TCS Research, IN"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2023,5,16]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.aei.2020.101043"},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2021.3074319"},{"key":"e_1_3_2_1_3_1","volume-title":"Proceedings of Machine Learning and Systems, I.\u00a0Dhillon, D.\u00a0Papailiopoulos, and V.\u00a0Sze (Eds.). Vol.\u00a02. mlsys.org, USA, 44\u201357","author":"Ahn Byung\u00a0Hoon","year":"2020","unstructured":"Byung\u00a0Hoon Ahn , Jinwon Lee , Jamie\u00a0Menjay Lin , Hsin-Pai Cheng , Jilei Hou , and Hadi Esmaeilzadeh . 2020 . Ordering Chaos: Memory-Aware Scheduling of Irregularly Wired Neural Networks for Edge Devices . In Proceedings of Machine Learning and Systems, I.\u00a0Dhillon, D.\u00a0Papailiopoulos, and V.\u00a0Sze (Eds.). Vol.\u00a02. mlsys.org, USA, 44\u201357 . https:\/\/proceedings.mlsys.org\/paper\/2020\/file\/9bf31c7ff062936a96d3c8bd1f8f2ff3-Paper.pdf Byung\u00a0Hoon Ahn, Jinwon Lee, Jamie\u00a0Menjay Lin, Hsin-Pai Cheng, Jilei Hou, and Hadi Esmaeilzadeh. 2020. Ordering Chaos: Memory-Aware Scheduling of Irregularly Wired Neural Networks for Edge Devices. In Proceedings of Machine Learning and Systems, I.\u00a0Dhillon, D.\u00a0Papailiopoulos, and V.\u00a0Sze (Eds.). Vol.\u00a02. mlsys.org, USA, 44\u201357. https:\/\/proceedings.mlsys.org\/paper\/2020\/file\/9bf31c7ff062936a96d3c8bd1f8f2ff3-Paper.pdf"},{"key":"e_1_3_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICOVET50258.2020.9230145"},{"key":"e_1_3_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/MIC50194.2020.9209615"},{"key":"e_1_3_2_1_7_1","volume-title":"Preemptive All-reduce Scheduling for Expediting Distributed DNN Training. In IEEE INFOCOM 2020 - IEEE Conference on Computer Communications. IEEE, NJ, 626\u2013635","author":"Bao Yixin","year":"2020","unstructured":"Yixin Bao , Yanghua Peng , Yangrui Chen , and Chuan Wu . 2020 . Preemptive All-reduce Scheduling for Expediting Distributed DNN Training. In IEEE INFOCOM 2020 - IEEE Conference on Computer Communications. IEEE, NJ, 626\u2013635 . Yixin Bao, Yanghua Peng, Yangrui Chen, and Chuan Wu. 2020. Preemptive All-reduce Scheduling for Expediting Distributed DNN Training. In IEEE INFOCOM 2020 - IEEE Conference on Computer Communications. IEEE, NJ, 626\u2013635."},{"key":"e_1_3_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.3390\/fi11040100"},{"key":"e_1_3_2_1_9_1","doi-asserted-by":"crossref","unstructured":"Chunlei Chen Peng Zhang Huixiang Zhang Jiangyan Dai Yugen Yi Huihui Zhang and Yonghui Zhang. 2020. Deep Learning on Computational-Resource-Limited Platforms: A Survey. Mob. Inf. Syst. 2020(2020) 8454327:1\u20138454327:19.  Chunlei Chen Peng Zhang Huixiang Zhang Jiangyan Dai Yugen Yi Huihui Zhang and Yonghui Zhang. 2020. Deep Learning on Computational-Resource-Limited Platforms: A Survey. Mob. Inf. Syst. 2020(2020) 8454327:1\u20138454327:19.","DOI":"10.1155\/2020\/8454327"},{"key":"e_1_3_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2021.3081353"},{"key":"e_1_3_2_1_11_1","unstructured":"Yoojin Choi Mostafa El-Khamy and Jungwon Lee. 2016. Towards the Limit of Network Quantization. CoRR abs\/1612.01543(2016). http:\/\/arxiv.org\/abs\/1612.01543  Yoojin Choi Mostafa El-Khamy and Jungwon Lee. 2016. Towards the Limit of Network Quantization. CoRR abs\/1612.01543(2016). http:\/\/arxiv.org\/abs\/1612.01543"},{"key":"e_1_3_2_1_12_1","unstructured":"Giulia Crocioni Giambattista Gruosso Danilo Pau Davide Denaro Luigi Zambrano and Giuseppe di Giore. 2021. Characterization of Neural Networks Automatically Mapped on Automotive-grade Microcontrollers. CoRR abs\/2103.00201(2021). https:\/\/arxiv.org\/abs\/2103.00201  Giulia Crocioni Giambattista Gruosso Danilo Pau Davide Denaro Luigi Zambrano and Giuseppe di Giore. 2021. Characterization of Neural Networks Automatically Mapped on Automotive-grade Microcontrollers. CoRR abs\/2103.00201(2021). https:\/\/arxiv.org\/abs\/2103.00201"},{"key":"e_1_3_2_1_13_1","volume-title":"Proceedings of Machine Learning and Systems, A.\u00a0Smola, A.\u00a0Dimakis, and I.\u00a0Stoica (Eds.). Vol.\u00a03. mlsys.org, USA, 800\u2013811","author":"David Robert","year":"2021","unstructured":"Robert David , Jared Duke , Advait Jain , Vijay Janapa\u00a0Reddi , Nat Jeffries , Jian Li , Nick Kreeger , Ian Nappier , Meghna Natraj , Tiezhen Wang , Pete Warden , and Rocky Rhodes . 2021 . TensorFlow Lite Micro: Embedded Machine Learning for TinyML Systems . In Proceedings of Machine Learning and Systems, A.\u00a0Smola, A.\u00a0Dimakis, and I.\u00a0Stoica (Eds.). Vol.\u00a03. mlsys.org, USA, 800\u2013811 . https:\/\/proceedings.mlsys.org\/paper\/2021\/file\/d2ddea18f00665ce8623e36bd4e3c7c5-Paper.pdf Robert David, Jared Duke, Advait Jain, Vijay Janapa\u00a0Reddi, Nat Jeffries, Jian Li, Nick Kreeger, Ian Nappier, Meghna Natraj, Tiezhen Wang, Pete Warden, and Rocky Rhodes. 2021. TensorFlow Lite Micro: Embedded Machine Learning for TinyML Systems. In Proceedings of Machine Learning and Systems, A.\u00a0Smola, A.\u00a0Dimakis, and I.\u00a0Stoica (Eds.). Vol.\u00a03. mlsys.org, USA, 800\u2013811. https:\/\/proceedings.mlsys.org\/paper\/2021\/file\/d2ddea18f00665ce8623e36bd4e3c7c5-Paper.pdf"},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1049\/et.2018.0529"},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1145\/3373376.3378473"},{"key":"e_1_3_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2009.5206848"},{"key":"#cr-split#-e_1_3_2_1_17_1.1","doi-asserted-by":"crossref","unstructured":"Swarnava Dey and Ranjan Dasgupta. 2009. Fast Boot User Experience Using Adaptive Storage Partitioning. In 2009 Computation World: Future Computing Service Computation Cognitive Adaptive Content Patterns. 113-118. https:\/\/doi.org\/10.1109\/ComputationWorld.2009.121 10.1109\/ComputationWorld.2009.121","DOI":"10.1109\/ComputationWorld.2009.121"},{"key":"#cr-split#-e_1_3_2_1_17_1.2","doi-asserted-by":"crossref","unstructured":"Swarnava Dey and Ranjan Dasgupta. 2009. Fast Boot User Experience Using Adaptive Storage Partitioning. In 2009 Computation World: Future Computing Service Computation Cognitive Adaptive Content Patterns. 113-118. https:\/\/doi.org\/10.1109\/ComputationWorld.2009.121","DOI":"10.1109\/ComputationWorld.2009.121"},{"key":"e_1_3_2_1_18_1","volume-title":"2019 IEEE International Conference on Pervasive Computing and Communications Workshops (PerCom Workshops). 855\u2013861","author":"Dey S.","unstructured":"S. Dey , J. Mondal , and A. Mukherjee . 2019. Offloaded Execution of Deep Learning Inference at Edge: Challenges and Insights . In 2019 IEEE International Conference on Pervasive Computing and Communications Workshops (PerCom Workshops). 855\u2013861 . S. Dey, J. Mondal, and A. Mukherjee. 2019. Offloaded Execution of Deep Learning Inference at Edge: Challenges and Insights. In 2019 IEEE International Conference on Pervasive Computing and Communications Workshops (PerCom Workshops). 855\u2013861."},{"key":"e_1_3_2_1_19_1","volume-title":"Implementing Deep Learning and Inferencing on Fog and Edge Computing Systems. 2018 IEEE International Conference on Pervasive Computing and Communications Workshops (PerCom Workshops)(2018)","author":"Dey Swarnava","year":"2018","unstructured":"Swarnava Dey and Arijit Mukherjee . 2018 . Implementing Deep Learning and Inferencing on Fog and Edge Computing Systems. 2018 IEEE International Conference on Pervasive Computing and Communications Workshops (PerCom Workshops)(2018) , 818\u2013823. Swarnava Dey and Arijit Mukherjee. 2018. Implementing Deep Learning and Inferencing on Fog and Edge Computing Systems. 2018 IEEE International Conference on Pervasive Computing and Communications Workshops (PerCom Workshops)(2018), 818\u2013823."},{"key":"e_1_3_2_1_20_1","volume-title":"Implementing Deep Learning and Inferencing on Fog and Edge Computing Systems. In 2018 IEEE International Conference on Pervasive Computing and Communications Workshops. IEEE, NJ, 818\u2013823","author":"Dey S.","unstructured":"S. Dey and A. Mukherjee . 2018 . Implementing Deep Learning and Inferencing on Fog and Edge Computing Systems. In 2018 IEEE International Conference on Pervasive Computing and Communications Workshops. IEEE, NJ, 818\u2013823 . S. Dey and A. Mukherjee. 2018. Implementing Deep Learning and Inferencing on Fog and Edge Computing Systems. In 2018 IEEE International Conference on Pervasive Computing and Communications Workshops. IEEE, NJ, 818\u2013823."},{"key":"e_1_3_2_1_21_1","volume-title":"Proceedings of the 1st ACM International Workshop on Smart Cities and Fog Computing","author":"Dey Swarnava","unstructured":"Swarnava Dey , Arijit Mukherjee , Arpan Pal , and P. Balamuralidhar . 2018. Partitioning of CNN Models for Execution on Fog Devices . In Proceedings of the 1st ACM International Workshop on Smart Cities and Fog Computing ( Shenzhen, China) (CitiFog\u201918). ACM, New York, NY, USA, 19\u201324. Swarnava Dey, Arijit Mukherjee, Arpan Pal, and P. Balamuralidhar. 2018. Partitioning of CNN Models for Execution on Fog Devices. In Proceedings of the 1st ACM International Workshop on Smart Cities and Fog Computing (Shenzhen, China) (CitiFog\u201918). ACM, New York, NY, USA, 19\u201324."},{"key":"e_1_3_2_1_22_1","volume-title":"FruitVegCNN: Power-and Memory-Efficient Classification of Fruits & Vegetables Using CNN in Mobile MPSoC. In 2020 IEEE 17th India Council International Conference (INDICON). IEEE, NJ, 1\u20137.","author":"Dey Somdip","year":"2020","unstructured":"Somdip Dey , Suman Saha , Amit Singh , and Klaus McDonald-Maier . 2020 . FruitVegCNN: Power-and Memory-Efficient Classification of Fruits & Vegetables Using CNN in Mobile MPSoC. In 2020 IEEE 17th India Council International Conference (INDICON). IEEE, NJ, 1\u20137. Somdip Dey, Suman Saha, Amit Singh, and Klaus McDonald-Maier. 2020. FruitVegCNN: Power-and Memory-Efficient Classification of Fruits & Vegetables Using CNN in Mobile MPSoC. In 2020 IEEE 17th India Council International Conference (INDICON). IEEE, NJ, 1\u20137."},{"key":"e_1_3_2_1_23_1","volume-title":"Automation & Test in Europe Conference & Exhibition (DATE). IEEE, NJ, 1728\u20131733","author":"Dey Somdip","year":"2020","unstructured":"Somdip Dey , Amit\u00a0Kumar Singh , Xiaohang Wang , and Klaus McDonald-Maier . 2020 . User interaction aware reinforcement learning for power and thermal efficiency of CPU-GPU mobile MPSoCs. In 2020 Design , Automation & Test in Europe Conference & Exhibition (DATE). IEEE, NJ, 1728\u20131733 . Somdip Dey, Amit\u00a0Kumar Singh, Xiaohang Wang, and Klaus McDonald-Maier. 2020. User interaction aware reinforcement learning for power and thermal efficiency of CPU-GPU mobile MPSoCs. In 2020 Design, Automation & Test in Europe Conference & Exhibition (DATE). IEEE, NJ, 1728\u20131733."},{"key":"e_1_3_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2020.2985932"},{"key":"e_1_3_2_1_25_1","unstructured":"Banbury et al.. 2021. MLPerf Tiny Benchmark. CoRR abs\/2106.07597(2021). arXiv:2106.07597https:\/\/arxiv.org\/abs\/2106.07597  Banbury et al.. 2021. MLPerf Tiny Benchmark. CoRR abs\/2106.07597(2021). arXiv:2106.07597https:\/\/arxiv.org\/abs\/2106.07597"},{"key":"e_1_3_2_1_26_1","volume-title":"Retrieved","author":"Franklin Dustin","year":"2018","unstructured":"Dustin Franklin . 2018 . NVIDIA Jetson AGX Xavier Delivers 32 TeraOps for New Era of AI in Robotics . Retrieved March 23, 2022 from https:\/\/developer.nvidia.com\/blog\/nvidia-jetson-agx-xavier-32-teraops-ai-robotics Dustin Franklin. 2018. NVIDIA Jetson AGX Xavier Delivers 32 TeraOps for New Era of AI in Robotics. Retrieved March 23, 2022 from https:\/\/developer.nvidia.com\/blog\/nvidia-jetson-agx-xavier-32-teraops-ai-robotics"},{"key":"e_1_3_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/AICAS48895.2020.9074001"},{"key":"e_1_3_2_1_28_1","volume-title":"Deep Compression: Compressing Deep Neural Network with Pruning, Trained Quantization and Huffman Coding. In 4th International Conference on Learning Representations, ICLR","author":"Han Song","year":"2016","unstructured":"Song Han , Huizi Mao , and William\u00a0 J. Dally . 2016 . Deep Compression: Compressing Deep Neural Network with Pruning, Trained Quantization and Huffman Coding. In 4th International Conference on Learning Representations, ICLR 2016, San Juan, Puerto Rico , May 2-4, 2016, Conference Track Proceedings, Yoshua Bengio and Yann LeCun (Eds.). ICLR, CA. Song Han, Huizi Mao, and William\u00a0J. Dally. 2016. Deep Compression: Compressing Deep Neural Network with Pruning, Trained Quantization and Huffman Coding. In 4th International Conference on Learning Representations, ICLR 2016, San Juan, Puerto Rico, May 2-4, 2016, Conference Track Proceedings, Yoshua Bengio and Yann LeCun (Eds.). ICLR, CA."},{"key":"e_1_3_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1109\/12.144628"},{"key":"e_1_3_2_1_30_1","volume-title":"Mobile AI 2021 Challenge: Report. In 2021 IEEE\/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW). IEEE, NJ, 2503\u20132514","author":"Ignatov Andrey","year":"2021","unstructured":"Andrey Ignatov , Cheng-Ming Chiang , and Hsien-Kai et. \u00a0al . Kuo. 2021 . Learned Smartphone ISP on Mobile NPUs with Deep Learning , Mobile AI 2021 Challenge: Report. In 2021 IEEE\/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW). IEEE, NJ, 2503\u20132514 . Andrey Ignatov, Cheng-Ming Chiang, and Hsien-Kai et.\u00a0al. Kuo. 2021. Learned Smartphone ISP on Mobile NPUs with Deep Learning, Mobile AI 2021 Challenge: Report. In 2021 IEEE\/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW). IEEE, NJ, 2503\u20132514."},{"key":"e_1_3_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.1145\/3267809.3267828"},{"key":"e_1_3_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1145\/3037697.3037698"},{"key":"e_1_3_2_1_33_1","unstructured":"Jong\u00a0Hwan Ko Taesik Na Mohammad\u00a0Faisal Amir and S. Mukhopadhyay. 2018. Edge-Host Partitioning of Deep Neural Networks with Feature Space Encoding for Resource-Constrained Internet-of-Things Platforms. In 2018 15th IEEE International Conference on Advanced Video and Signal Based Surveillance (AVSS). IEEE NJ 1\u20136.  Jong\u00a0Hwan Ko Taesik Na Mohammad\u00a0Faisal Amir and S. Mukhopadhyay. 2018. Edge-Host Partitioning of Deep Neural Networks with Feature Space Encoding for Resource-Constrained Internet-of-Things Platforms. In 2018 15th IEEE International Conference on Advanced Video and Signal Based Surveillance (AVSS). IEEE NJ 1\u20136."},{"key":"e_1_3_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/VLSI-DAT.2018.8373244"},{"key":"e_1_3_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISSCC19947.2020.9063111"},{"key":"e_1_3_2_1_37_1","volume-title":"MCUNet: Tiny Deep Learning on IoT Devices. In Advances in Neural Information Processing Systems 33: Annual Conference on Neural Information Processing Systems 2020","author":"Lin Ji","year":"2020","unstructured":"Ji Lin , Wei-Ming Chen , Yujun Lin , John Cohn , Chuang Gan , and Song Han . 2020 . MCUNet: Tiny Deep Learning on IoT Devices. In Advances in Neural Information Processing Systems 33: Annual Conference on Neural Information Processing Systems 2020 , NeurIPS 2020, December 6-12, 2020, virtual, Hugo Larochelle, Marc\u2019Aurelio Ranzato, Raia Hadsell, Maria-Florina Balcan, and Hsuan-Tien Lin (Eds.). NeurIPS, Canada. Ji Lin, Wei-Ming Chen, Yujun Lin, John Cohn, Chuang Gan, and Song Han. 2020. MCUNet: Tiny Deep Learning on IoT Devices. In Advances in Neural Information Processing Systems 33: Annual Conference on Neural Information Processing Systems 2020, NeurIPS 2020, December 6-12, 2020, virtual, Hugo Larochelle, Marc\u2019Aurelio Ranzato, Raia Hadsell, Maria-Florina Balcan, and Hsuan-Tien Lin (Eds.). NeurIPS, Canada."},{"key":"e_1_3_2_1_38_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCAD.2017.8203852"},{"key":"e_1_3_2_1_39_1","volume-title":"Partitioning Convolutional Neural Networks to Maximize the Inference Rate on Constrained IoT Devices. Future Internet 11, 10","author":"Campos\u00a0de Oliveira Fab\u00edola Martins","year":"2019","unstructured":"Fab\u00edola Martins Campos\u00a0de Oliveira and Edson Borin . 2019. Partitioning Convolutional Neural Networks to Maximize the Inference Rate on Constrained IoT Devices. Future Internet 11, 10 ( 2019 ). https:\/\/doi.org\/10.3390\/fi11100209 10.3390\/fi11100209 Fab\u00edola Martins Campos\u00a0de Oliveira and Edson Borin. 2019. Partitioning Convolutional Neural Networks to Maximize the Inference Rate on Constrained IoT Devices. Future Internet 11, 10 (2019). https:\/\/doi.org\/10.3390\/fi11100209"},{"key":"e_1_3_2_1_40_1","unstructured":"Mirgahney Mohamed Gabriele Cesa Taco\u00a0S. Cohen and Max Welling. 2020. A Data and Compute Efficient Design for Limited-Resources Deep Learning. CoRR abs\/2004.09691(2020). https:\/\/arxiv.org\/abs\/2004.09691  Mirgahney Mohamed Gabriele Cesa Taco\u00a0S. Cohen and Max Welling. 2020. A Data and Compute Efficient Design for Limited-Resources Deep Learning. CoRR abs\/2004.09691(2020). https:\/\/arxiv.org\/abs\/2004.09691"},{"key":"e_1_3_2_1_41_1","volume-title":"Accelerated Fire Detection and Localization at Edge. ACM Trans. Embed. Comput. Syst. (dec","author":"Mukherjee Arijit","year":"2022","unstructured":"Arijit Mukherjee , Jayeeta Mondal , and Swarnava Dey . 2022. Accelerated Fire Detection and Localization at Edge. ACM Trans. Embed. Comput. Syst. (dec 2022 ). https:\/\/doi.org\/10.1145\/3510027 Just Accepted . 10.1145\/3510027 Arijit Mukherjee, Jayeeta Mondal, and Swarnava Dey. 2022. Accelerated Fire Detection and Localization at Edge. ACM Trans. Embed. Comput. Syst. (dec 2022). https:\/\/doi.org\/10.1145\/3510027 Just Accepted."},{"key":"e_1_3_2_1_42_1","volume-title":"Retrieved","author":"Network Qualcomm\u00a0Developer","year":"2021","unstructured":"Qualcomm\u00a0Developer Network . 2021 . Qualcomm Neural Processing SDK for AI . Retrieved July 14, 2021 from https:\/\/developer.qualcomm.com\/software\/qualcomm-neural-processing-sdk Qualcomm\u00a0Developer Network. 2021. Qualcomm Neural Processing SDK for AI. Retrieved July 14, 2021 from https:\/\/developer.qualcomm.com\/software\/qualcomm-neural-processing-sdk"},{"key":"e_1_3_2_1_43_1","volume-title":"Retrieved","author":"Newsroom Samsung","year":"2018","unstructured":"Samsung Newsroom . 2018 . Samsung Optimizes Premium Exynos 9 Series 9810 for AI Applications and Richer Multimedia Content . Retrieved July 14, 2021 from https:\/\/news.samsung.com\/global\/samsung-optimizes-premium-exynos-9-series-9810-for-ai-applications-and-richer-multimedia-content Samsung Newsroom. 2018. Samsung Optimizes Premium Exynos 9 Series 9810 for AI Applications and Richer Multimedia Content. Retrieved July 14, 2021 from https:\/\/news.samsung.com\/global\/samsung-optimizes-premium-exynos-9-series-9810-for-ai-applications-and-richer-multimedia-content"},{"key":"e_1_3_2_1_44_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCE-Asia49877.2020.9277463"},{"key":"e_1_3_2_1_45_1","volume-title":"Future of Healthcare\u2014Sensor Data-Driven Prognosis","author":"Pal Arpan","unstructured":"Arpan Pal , Arijit Mukherjee , and Swarnava Dey . 2016. Future of Healthcare\u2014Sensor Data-Driven Prognosis . Springer International Publishing , Cham , 93\u2013109. https:\/\/doi.org\/10.1007\/978-3-319-42141-4_9 10.1007\/978-3-319-42141-4_9 Arpan Pal, Arijit Mukherjee, and Swarnava Dey. 2016. Future of Healthcare\u2014Sensor Data-Driven Prognosis. Springer International Publishing, Cham, 93\u2013109. https:\/\/doi.org\/10.1007\/978-3-319-42141-4_9"},{"key":"e_1_3_2_1_46_1","volume-title":"Comparing Industry Frameworks with Deeply Quantized Neural Networks on Microcontrollers. In 2021 IEEE International Conference on Consumer Electronics (ICCE). IEEE, NJ, 1\u20136.","author":"Pau Danilo","year":"2021","unstructured":"Danilo Pau , Marco Lattuada , Francesco Loro , Antonio De\u00a0Vita , and Gian Domenico\u00a0Licciardo . 2021 . Comparing Industry Frameworks with Deeply Quantized Neural Networks on Microcontrollers. In 2021 IEEE International Conference on Consumer Electronics (ICCE). IEEE, NJ, 1\u20136. Danilo Pau, Marco Lattuada, Francesco Loro, Antonio De\u00a0Vita, and Gian Domenico\u00a0Licciardo. 2021. Comparing Industry Frameworks with Deeply Quantized Neural Networks on Microcontrollers. In 2021 IEEE International Conference on Consumer Electronics (ICCE). IEEE, NJ, 1\u20136."},{"key":"e_1_3_2_1_47_1","volume-title":"Electronics and Mechatronics Conference (IEMTRONICS). IEEE, NJ, 1\u20134.","author":"Prasad P\u00a0Kavyashree","year":"2021","unstructured":"S P\u00a0Kavyashree Prasad and Mohamed El-Sharkawy . 2021 . Deployment of Compressed MobileNet V3 on iMX RT 1060. In 2021 IEEE International IOT , Electronics and Mechatronics Conference (IEMTRONICS). IEEE, NJ, 1\u20134. SP\u00a0Kavyashree Prasad and Mohamed El-Sharkawy. 2021. Deployment of Compressed MobileNet V3 on iMX RT 1060. In 2021 IEEE International IOT, Electronics and Mechatronics Conference (IEMTRONICS). IEEE, NJ, 1\u20134."},{"key":"e_1_3_2_1_48_1","doi-asserted-by":"publisher","DOI":"10.1145\/3097895.3097903"},{"key":"e_1_3_2_1_49_1","volume-title":"\u00a0Corchado Rodr\u00edguez","author":"S\u00e1nchez Sergio\u00a0M\u00e1rquez","year":"2020","unstructured":"Sergio\u00a0M\u00e1rquez S\u00e1nchez , Francisco Lecumberri , Vishwani Sati , Ashish Arora , Niloufar Shoeibi , Sara Rodr\u00edguez , and Juan M . \u00a0Corchado Rodr\u00edguez . 2020 . Edge Computing Driven Smart Personal Protective System Deployed on NVIDIA Jetson and Integrated with ROS. In Highlights in Practical Applications of Agents, Multi-Agent Systems, and Trust-worthiness. The PAAMS Collection, Fernando De\u00a0La\u00a0Prieta, Philippe Mathieu, Jaime\u00a0Andr\u00e9s Rinc\u00f3n\u00a0Arango, Alia El\u00a0Bolock, Elena Del\u00a0Val, Jaume Jord\u00e1n\u00a0Prunera, Jo\u00e3o Carneiro, Rub\u00e9n Fuentes, Fernando Lopes, and Vicente Julian (Eds.). Springer International Publishing , Cham, 385\u2013393. Sergio\u00a0M\u00e1rquez S\u00e1nchez, Francisco Lecumberri, Vishwani Sati, Ashish Arora, Niloufar Shoeibi, Sara Rodr\u00edguez, and Juan M.\u00a0Corchado Rodr\u00edguez. 2020. Edge Computing Driven Smart Personal Protective System Deployed on NVIDIA Jetson and Integrated with ROS. In Highlights in Practical Applications of Agents, Multi-Agent Systems, and Trust-worthiness. The PAAMS Collection, Fernando De\u00a0La\u00a0Prieta, Philippe Mathieu, Jaime\u00a0Andr\u00e9s Rinc\u00f3n\u00a0Arango, Alia El\u00a0Bolock, Elena Del\u00a0Val, Jaume Jord\u00e1n\u00a0Prunera, Jo\u00e3o Carneiro, Rub\u00e9n Fuentes, Fernando Lopes, and Vicente Julian (Eds.). Springer International Publishing, Cham, 385\u2013393."},{"key":"e_1_3_2_1_50_1","volume-title":"Retrieved","author":"Sharma Jairam","year":"2022","unstructured":"Jairam Sharma . 2022 . Building Industrial embedded deep learning inference pipelines with TensorRT . Retrieved March 23, 2022 from https:\/\/learnopencv.com\/building-industrial-embedded-deep-learning-inference-pipelines-with-tensorrt\/ Jairam Sharma. 2022. Building Industrial embedded deep learning inference pipelines with TensorRT. Retrieved March 23, 2022 from https:\/\/learnopencv.com\/building-industrial-embedded-deep-learning-inference-pipelines-with-tensorrt\/"},{"key":"e_1_3_2_1_51_1","unstructured":"Duncan Stewart Jeff Loucks Mark Casey and Craig Wigginton. 2019. Bringing AI to the device: Edge AI chips come into their own. https:\/\/www2.deloitte.com\/us\/en\/insights\/industry\/technology\/technology-media-and-telecom-predictions\/2020\/ai-chips.html.  Duncan Stewart Jeff Loucks Mark Casey and Craig Wigginton. 2019. Bringing AI to the device: Edge AI chips come into their own. https:\/\/www2.deloitte.com\/us\/en\/insights\/industry\/technology\/technology-media-and-telecom-predictions\/2020\/ai-chips.html."},{"key":"e_1_3_2_1_52_1","volume-title":"2017 IEEE 37th International Conference on Distributed Computing Systems (ICDCS). IEEE, NJ, 328\u2013339","author":"Teerapittayanon Surat","unstructured":"Surat Teerapittayanon , Bradley McDanel , and H.T. Kung . 2017. Distributed Deep Neural Networks Over the Cloud, the Edge and End Devices . In 2017 IEEE 37th International Conference on Distributed Computing Systems (ICDCS). IEEE, NJ, 328\u2013339 . Surat Teerapittayanon, Bradley McDanel, and H.T. Kung. 2017. Distributed Deep Neural Networks Over the Cloud, the Edge and End Devices. In 2017 IEEE 37th International Conference on Distributed Computing Systems (ICDCS). IEEE, NJ, 328\u2013339."},{"key":"e_1_3_2_1_53_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDCS.2017.226"},{"key":"e_1_3_2_1_54_1","unstructured":"Rob van\u00a0der Meulen. 2018. What Edge Computing Means for Infrastructure and Operations Leaders. https:\/\/www.gartner.com\/smarterwithgartner\/what-edge-computing-means-for-infrastructure-and-operations-leaders.  Rob van\u00a0der Meulen. 2018. What Edge Computing Means for Infrastructure and Operations Leaders. https:\/\/www.gartner.com\/smarterwithgartner\/what-edge-computing-means-for-infrastructure-and-operations-leaders."},{"key":"e_1_3_2_1_55_1","doi-asserted-by":"publisher","DOI":"10.1186\/s13677-020-00172-z"},{"key":"e_1_3_2_1_56_1","doi-asserted-by":"publisher","DOI":"10.1080\/21642583.2018.1477634"},{"key":"e_1_3_2_1_57_1","volume-title":"Pipelined Data-Parallel CPU\/GPU Scheduling for Multi-DNN Real-Time Inference. In 2019 IEEE Real-Time Systems Symposium (RTSS). IEEE, NJ, 392\u2013405","author":"Xiang Yecheng","year":"2019","unstructured":"Yecheng Xiang and Hyoseung Kim . 2019 . Pipelined Data-Parallel CPU\/GPU Scheduling for Multi-DNN Real-Time Inference. In 2019 IEEE Real-Time Systems Symposium (RTSS). IEEE, NJ, 392\u2013405 . Yecheng Xiang and Hyoseung Kim. 2019. Pipelined Data-Parallel CPU\/GPU Scheduling for Multi-DNN Real-Time Inference. In 2019 IEEE Real-Time Systems Symposium (RTSS). IEEE, NJ, 392\u2013405."},{"key":"e_1_3_2_1_58_1","doi-asserted-by":"crossref","unstructured":"Y. Xing P. Kirkland G. Di\u00a0Caterina J. Soraghan and G. Matich. 2018. Real-Time Embedded Intelligence System: Emotion Recognition on Raspberry Pi with Intel NCS. In Artificial Neural Networks and Machine Learning \u2013 ICANN 2018 V\u011bra K\u016frkov\u00e1 Yannis Manolopoulos Barbara Hammer Lazaros Iliadis and Ilias Maglogiannis (Eds.). Springer International Publishing Cham 801\u2013808.  Y. Xing P. Kirkland G. Di\u00a0Caterina J. Soraghan and G. Matich. 2018. Real-Time Embedded Intelligence System: Emotion Recognition on Raspberry Pi with Intel NCS. In Artificial Neural Networks and Machine Learning \u2013 ICANN 2018 V\u011bra K\u016frkov\u00e1 Yannis Manolopoulos Barbara Hammer Lazaros Iliadis and Ilias Maglogiannis (Eds.). Springer International Publishing Cham 801\u2013808.","DOI":"10.1007\/978-3-030-01418-6_78"},{"key":"e_1_3_2_1_59_1","doi-asserted-by":"publisher","DOI":"10.1109\/MPRV.2021.3052532"},{"key":"e_1_3_2_1_60_1","unstructured":"Dingcheng Yang Wenjian Yu Ao Zhou Haoyuan Mu Gary Yao and Xiaoyi Wang. 2020. DP-Net: Dynamic Programming Guided Deep Neural Network Compression. CoRR abs\/2003.09615(2020).  Dingcheng Yang Wenjian Yu Ao Zhou Haoyuan Mu Gary Yao and Xiaoyi Wang. 2020. DP-Net: Dynamic Programming Guided Deep Neural Network Compression. CoRR abs\/2003.09615(2020)."},{"key":"e_1_3_2_1_61_1","doi-asserted-by":"publisher","DOI":"10.1145\/3384419.3430898"},{"key":"e_1_3_2_1_62_1","doi-asserted-by":"publisher","DOI":"10.1145\/3274783.3274840"},{"key":"e_1_3_2_1_63_1","doi-asserted-by":"publisher","DOI":"10.1145\/3318216.3363312"},{"key":"e_1_3_2_1_64_1","volume-title":"2nd USENIX Workshop on Hot Topics in Edge Computing, HotEdge 2019","author":"Zhou Li","year":"2019","unstructured":"Li Zhou , Hao Wen , Radu Teodorescu , and David H . \u00a0C. Du. 2019. Distributing Deep Neural Networks with Containerized Partitions at the Edge . In 2nd USENIX Workshop on Hot Topics in Edge Computing, HotEdge 2019 , Renton, WA, USA , July 9, 2019 , Irfan Ahmad and Swaminathan Sundararaman (Eds.). USENIX Association. https:\/\/www.usenix.org\/conference\/hotedge19\/presentation\/zhou Li Zhou, Hao Wen, Radu Teodorescu, and David H.\u00a0C. Du. 2019. Distributing Deep Neural Networks with Containerized Partitions at the Edge. In 2nd USENIX Workshop on Hot Topics in Edge Computing, HotEdge 2019, Renton, WA, USA, July 9, 2019, Irfan Ahmad and Swaminathan Sundararaman (Eds.). USENIX Association. https:\/\/www.usenix.org\/conference\/hotedge19\/presentation\/zhou"}],"event":{"name":"AIMLSystems 2022: The Second International Conference on AI-ML Systems","location":"Bangalore India","acronym":"AIMLSystems 2022"},"container-title":["Proceedings of the Second International Conference on AI-ML Systems"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3564121.3564796","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3564121.3564796","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,19]],"date-time":"2025-06-19T01:10:33Z","timestamp":1750295433000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3564121.3564796"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,10,12]]},"references-count":63,"alternative-id":["10.1145\/3564121.3564796","10.1145\/3564121"],"URL":"https:\/\/doi.org\/10.1145\/3564121.3564796","relation":{},"subject":[],"published":{"date-parts":[[2022,10,12]]},"assertion":[{"value":"2023-05-16","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}