{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,19]],"date-time":"2026-06-19T16:35:38Z","timestamp":1781886938691,"version":"3.54.5"},"publisher-location":"New York, NY, USA","reference-count":42,"publisher":"ACM","license":[{"start":{"date-parts":[[2023,5,9]],"date-time":"2023-05-09T00:00:00Z","timestamp":1683590400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2023,5,9]]},"DOI":"10.1145\/3583120.3586953","type":"proceedings-article","created":{"date-parts":[[2023,5,5]],"date-time":"2023-05-05T16:23:44Z","timestamp":1683303824000},"page":"1-1","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":14,"title":["POS: An Operator Scheduling Framework for Multi-model Inference on Edge Intelligent Computing"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0000-0003-2539-8257","authenticated-orcid":false,"given":"Ziyang","family":"Zhang","sequence":"first","affiliation":[{"name":"Harbin Institute of Technology, Harbin, China, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0009-7350-5361","authenticated-orcid":false,"given":"Huan","family":"Li","sequence":"additional","affiliation":[{"name":"Harbin Institute of Technology, Shenzhen, China, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-2080-1270","authenticated-orcid":false,"given":"Yang","family":"Zhao","sequence":"additional","affiliation":[{"name":"Harbin Institute of Technology, Shenzhen, China, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-6805-2649","authenticated-orcid":false,"given":"Changyao","family":"Lin","sequence":"additional","affiliation":[{"name":"Harbin Institute of Technology, Harbin, China, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-6209-6886","authenticated-orcid":false,"given":"Jie","family":"Liu","sequence":"additional","affiliation":[{"name":"Harbin Institute of Technology, Shenzhen, China, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2023,5,9]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/JPROC.2019.2921977"},{"key":"e_1_3_2_1_2_1","volume-title":"13th USENIX Symposium on Operating Systems Design and Implementation (OSDI 18)","author":"Chen Tianqi","year":"2018","unstructured":"Tianqi Chen , Thierry Moreau , Ziheng Jiang , Lianmin Zheng , Eddie Yan , Haichen Shen , Meghan Cowan , Leyuan Wang , Yuwei Hu , Luis Ceze , 2018 . { TVM} : An automated { End-to-End} optimizing compiler for deep learning . In 13th USENIX Symposium on Operating Systems Design and Implementation (OSDI 18) . 578\u2013594. Tianqi Chen, Thierry Moreau, Ziheng Jiang, Lianmin Zheng, Eddie Yan, Haichen Shen, Meghan Cowan, Leyuan Wang, Yuwei Hu, Luis Ceze, 2018. { TVM} : An automated { End-to-End} optimizing compiler for deep learning. In 13th USENIX Symposium on Operating Systems Design and Implementation (OSDI 18). 578\u2013594."},{"key":"e_1_3_2_1_3_1","volume-title":"Learning to optimize tensor programs. Advances in Neural Information Processing Systems 31","author":"Chen Tianqi","year":"2018","unstructured":"Tianqi Chen , Lianmin Zheng , Eddie Yan , Ziheng Jiang , Thierry Moreau , Luis Ceze , Carlos Guestrin , and Arvind Krishnamurthy . 2018. Learning to optimize tensor programs. Advances in Neural Information Processing Systems 31 ( 2018 ). Tianqi Chen, Lianmin Zheng, Eddie Yan, Ziheng Jiang, Thierry Moreau, Luis Ceze, Carlos Guestrin, and Arvind Krishnamurthy. 2018. Learning to optimize tensor programs. Advances in Neural Information Processing Systems 31 (2018)."},{"key":"e_1_3_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1145\/3447548.3467132"},{"key":"e_1_3_2_1_5_1","volume-title":"Soft actor-critic for discrete action settings. arXiv preprint arXiv:1910.07207","author":"Christodoulou Petros","year":"2019","unstructured":"Petros Christodoulou . 2019. Soft actor-critic for discrete action settings. arXiv preprint arXiv:1910.07207 ( 2019 ). Petros Christodoulou. 2019. Soft actor-critic for discrete action settings. arXiv preprint arXiv:1910.07207 (2019)."},{"key":"e_1_3_2_1_6_1","volume-title":"Bert: Pre-training of deep bidirectional transformers for language understanding. arXiv preprint arXiv:1810.04805","author":"Devlin Jacob","year":"2018","unstructured":"Jacob Devlin , Ming-Wei Chang , Kenton Lee , and Kristina Toutanova . 2018 . Bert: Pre-training of deep bidirectional transformers for language understanding. arXiv preprint arXiv:1810.04805 (2018). Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018. Bert: Pre-training of deep bidirectional transformers for language understanding. arXiv preprint arXiv:1810.04805 (2018)."},{"key":"e_1_3_2_1_7_1","first-page":"167","article-title":"Ios: Inter-operator scheduler for cnn acceleration","volume":"3","author":"Ding Yaoyao","year":"2021","unstructured":"Yaoyao Ding , Ligeng Zhu , Zhihao Jia , Gennady Pekhimenko , and Song Han . 2021 . Ios: Inter-operator scheduler for cnn acceleration . Proceedings of Machine Learning and Systems 3 (2021), 167 \u2013 180 . Yaoyao Ding, Ligeng Zhu, Zhihao Jia, Gennady Pekhimenko, and Song Han. 2021. Ios: Inter-operator scheduler for cnn acceleration. Proceedings of Machine Learning and Systems 3 (2021), 167\u2013180.","journal-title":"Proceedings of Machine Learning and Systems"},{"key":"e_1_3_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2012.6248074"},{"key":"e_1_3_2_1_9_1","volume-title":"Soft actor-critic algorithms and applications. arXiv preprint arXiv:1812.05905","author":"Haarnoja Tuomas","year":"2018","unstructured":"Tuomas Haarnoja , Aurick Zhou , Kristian Hartikainen , George Tucker , Sehoon Ha , Jie Tan , Vikash Kumar , Henry Zhu , Abhishek Gupta , Pieter Abbeel , 2018. Soft actor-critic algorithms and applications. arXiv preprint arXiv:1812.05905 ( 2018 ). Tuomas Haarnoja, Aurick Zhou, Kristian Hartikainen, George Tucker, Sehoon Ha, Jie Tan, Vikash Kumar, Henry Zhu, Abhishek Gupta, Pieter Abbeel, 2018. Soft actor-critic algorithms and applications. arXiv preprint arXiv:1812.05905 (2018)."},{"key":"e_1_3_2_1_10_1","volume-title":"16th USENIX Symposium on Operating Systems Design and Implementation (OSDI 22)","author":"Han Mingcong","year":"2022","unstructured":"Mingcong Han , Hanze Zhang , Rong Chen , and Haibo Chen . 2022 . Microsecond-scale Preemption for Concurrent { GPU-accelerated}{ DNN} Inferences . In 16th USENIX Symposium on Operating Systems Design and Implementation (OSDI 22) . 539\u2013558. Mingcong Han, Hanze Zhang, Rong Chen, and Haibo Chen. 2022. Microsecond-scale Preemption for Concurrent { GPU-accelerated}{ DNN} Inferences. In 16th USENIX Symposium on Operating Systems Design and Implementation (OSDI 22). 539\u2013558."},{"key":"e_1_3_2_1_11_1","volume-title":"Deep compression: Compressing deep neural networks with pruning, trained quantization and huffman coding. arXiv preprint arXiv:1510.00149","author":"Han Song","year":"2015","unstructured":"Song Han , Huizi Mao , and William\u00a0 J Dally . 2015. Deep compression: Compressing deep neural networks with pruning, trained quantization and huffman coding. arXiv preprint arXiv:1510.00149 ( 2015 ). Song Han, Huizi Mao, and William\u00a0J Dally. 2015. Deep compression: Compressing deep neural networks with pruning, trained quantization and huffman coding. arXiv preprint arXiv:1510.00149 (2015)."},{"key":"e_1_3_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.90"},{"key":"e_1_3_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00140"},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPWRD.2020.3024965"},{"key":"e_1_3_2_1_15_1","volume-title":"Highly accurate protein structure prediction with AlphaFold. Nature 596, 7873","author":"Jumper John","year":"2021","unstructured":"John Jumper , Richard Evans , Alexander Pritzel , Tim Green , Michael Figurnov , Olaf Ronneberger , Kathryn Tunyasuvunakool , Russ Bates , Augustin \u017d\u00eddek , Anna Potapenko , 2021. Highly accurate protein structure prediction with AlphaFold. Nature 596, 7873 ( 2021 ), 583\u2013589. John Jumper, Richard Evans, Alexander Pritzel, Tim Green, Michael Figurnov, Olaf Ronneberger, Kathryn Tunyasuvunakool, Russ Bates, Augustin \u017d\u00eddek, Anna Potapenko, 2021. Highly accurate protein structure prediction with AlphaFold. Nature 596, 7873 (2021), 583\u2013589."},{"key":"e_1_3_2_1_16_1","volume-title":"Key points estimation and point instance segmentation approach for lane detection","author":"Ko Yeongmin","year":"2021","unstructured":"Yeongmin Ko , Younkwan Lee , Shoaib Azam , Farzeen Munir , Moongu Jeon , and Witold Pedrycz . 2021. Key points estimation and point instance segmentation approach for lane detection . IEEE Transactions on Intelligent Transportation Systems ( 2021 ). Yeongmin Ko, Younkwan Lee, Shoaib Azam, Farzeen Munir, Moongu Jeon, and Witold Pedrycz. 2021. Key points estimation and point instance segmentation approach for lane detection. IEEE Transactions on Intelligent Transportation Systems (2021)."},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/TITS.2021.3095161"},{"key":"e_1_3_2_1_18_1","first-page":"8343","article-title":"Nimble: Lightweight and parallel gpu task scheduling for deep learning","volume":"33","author":"Kwon Woosuk","year":"2020","unstructured":"Woosuk Kwon , Gyeong-In Yu , Eunji Jeong , and Byung-Gon Chun . 2020 . Nimble: Lightweight and parallel gpu task scheduling for deep learning . Advances in Neural Information Processing Systems 33 (2020), 8343 \u2013 8354 . Woosuk Kwon, Gyeong-In Yu, Eunji Jeong, and Byung-Gon Chun. 2020. Nimble: Lightweight and parallel gpu task scheduling for deep learning. Advances in Neural Information Processing Systems 33 (2020), 8343\u20138354.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1145\/3372224.3419194"},{"key":"e_1_3_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-46448-0_8"},{"key":"e_1_3_2_1_21_1","volume-title":"Continuous control with deep reinforcement learning. arXiv preprint arXiv:1509.02971","author":"Lillicrap P","year":"2015","unstructured":"Timothy\u00a0 P Lillicrap , Jonathan\u00a0 J Hunt , Alexander Pritzel , Nicolas Heess , Tom Erez , Yuval Tassa , David Silver , and Daan Wierstra . 2015. Continuous control with deep reinforcement learning. arXiv preprint arXiv:1509.02971 ( 2015 ). Timothy\u00a0P Lillicrap, Jonathan\u00a0J Hunt, Alexander Pritzel, Nicolas Heess, Tom Erez, Yuval Tassa, David Silver, and Daan Wierstra. 2015. Continuous control with deep reinforcement learning. arXiv preprint arXiv:1509.02971 (2015)."},{"key":"e_1_3_2_1_22_1","volume-title":"14th USENIX Symposium on Operating Systems Design and Implementation (OSDI 20)","author":"Ma Lingxiao","year":"2020","unstructured":"Lingxiao Ma , Zhiqiang Xie , Zhi Yang , Jilong Xue , Youshan Miao , Wei Cui , Wenxiang Hu , Fan Yang , Lintao Zhang , and Lidong Zhou . 2020 . Rammer: Enabling Holistic Deep Learning Compiler Optimizations with { rTasks} . In 14th USENIX Symposium on Operating Systems Design and Implementation (OSDI 20) . 881\u2013897. Lingxiao Ma, Zhiqiang Xie, Zhi Yang, Jilong Xue, Youshan Miao, Wei Cui, Wenxiang Hu, Fan Yang, Lintao Zhang, and Lidong Zhou. 2020. Rammer: Enabling Holistic Deep Learning Compiler Optimizations with { rTasks}. In 14th USENIX Symposium on Operating Systems Design and Implementation (OSDI 20). 881\u2013897."},{"key":"e_1_3_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1145\/3453483.3454083"},{"key":"e_1_3_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.00051"},{"key":"e_1_3_2_1_25_1","volume-title":"Mastering atari, go, chess and shogi by planning with a learned model. Nature 588, 7839","author":"Schrittwieser Julian","year":"2020","unstructured":"Julian Schrittwieser , Ioannis Antonoglou , Thomas Hubert , Karen Simonyan , Laurent Sifre , Simon Schmitt , Arthur Guez , Edward Lockhart , Demis Hassabis , Thore Graepel , 2020. Mastering atari, go, chess and shogi by planning with a learned model. Nature 588, 7839 ( 2020 ), 604\u2013609. Julian Schrittwieser, Ioannis Antonoglou, Thomas Hubert, Karen Simonyan, Laurent Sifre, Simon Schmitt, Arthur Guez, Edward Lockhart, Demis Hassabis, Thore Graepel, 2020. Mastering atari, go, chess and shogi by planning with a learned model. Nature 588, 7839 (2020), 604\u2013609."},{"key":"e_1_3_2_1_26_1","volume-title":"International conference on machine learning. PMLR","author":"Schulman John","year":"2015","unstructured":"John Schulman , Sergey Levine , Pieter Abbeel , Michael Jordan , and Philipp Moritz . 2015 . Trust region policy optimization . In International conference on machine learning. PMLR , 1889\u20131897. John Schulman, Sergey Levine, Pieter Abbeel, Michael Jordan, and Philipp Moritz. 2015. Trust region policy optimization. In International conference on machine learning. PMLR, 1889\u20131897."},{"key":"e_1_3_2_1_27_1","volume-title":"Proximal policy optimization algorithms. arXiv preprint arXiv:1707.06347","author":"Schulman John","year":"2017","unstructured":"John Schulman , Filip Wolski , Prafulla Dhariwal , Alec Radford , and Oleg Klimov . 2017. Proximal policy optimization algorithms. arXiv preprint arXiv:1707.06347 ( 2017 ). John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov. 2017. Proximal policy optimization algorithms. arXiv preprint arXiv:1707.06347 (2017)."},{"key":"e_1_3_2_1_28_1","first-page":"208","article-title":"Nimble: Efficiently compiling dynamic neural networks for model inference","volume":"3","author":"Shen Haichen","year":"2021","unstructured":"Haichen Shen , Jared Roesch , Zhi Chen , Wei Chen , Yong Wu , Mu Li , Vin Sharma , Zachary Tatlock , and Yida Wang . 2021 . Nimble: Efficiently compiling dynamic neural networks for model inference . Proceedings of Machine Learning and Systems 3 (2021), 208 \u2013 222 . Haichen Shen, Jared Roesch, Zhi Chen, Wei Chen, Yong Wu, Mu Li, Vin Sharma, Zachary Tatlock, and Yida Wang. 2021. Nimble: Efficiently compiling dynamic neural networks for model inference. Proceedings of Machine Learning and Systems 3 (2021), 208\u2013222.","journal-title":"Proceedings of Machine Learning and Systems"},{"key":"e_1_3_2_1_29_1","volume-title":"Edge computing: Vision and challenges","author":"Shi Weisong","year":"2016","unstructured":"Weisong Shi , Jie Cao , Quan Zhang , Youhuizi Li , and Lanyu Xu. 2016. Edge computing: Vision and challenges . IEEE internet of things journal 3, 5 ( 2016 ), 637\u2013646. Weisong Shi, Jie Cao, Quan Zhang, Youhuizi Li, and Lanyu Xu. 2016. Edge computing: Vision and challenges. IEEE internet of things journal 3, 5 (2016), 637\u2013646."},{"key":"e_1_3_2_1_30_1","volume-title":"Sequence to sequence learning with neural networks. Advances in neural information processing systems 27","author":"Sutskever Ilya","year":"2014","unstructured":"Ilya Sutskever , Oriol Vinyals , and Quoc\u00a0 V Le. 2014. Sequence to sequence learning with neural networks. Advances in neural information processing systems 27 ( 2014 ). Ilya Sutskever, Oriol Vinyals, and Quoc\u00a0V Le. 2014. Sequence to sequence learning with neural networks. Advances in neural information processing systems 27 (2014)."},{"key":"e_1_3_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.308"},{"key":"e_1_3_2_1_32_1","volume-title":"Improved semantic representations from tree-structured long short-term memory networks. arXiv preprint arXiv:1503.00075","author":"Tai Kai\u00a0Sheng","year":"2015","unstructured":"Kai\u00a0Sheng Tai , Richard Socher , and Christopher\u00a0 D Manning . 2015. Improved semantic representations from tree-structured long short-term memory networks. arXiv preprint arXiv:1503.00075 ( 2015 ). Kai\u00a0Sheng Tai, Richard Socher, and Christopher\u00a0D Manning. 2015. Improved semantic representations from tree-structured long short-term memory networks. arXiv preprint arXiv:1503.00075 (2015)."},{"key":"e_1_3_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.1145\/3447993.3448625"},{"key":"e_1_3_2_1_34_1","first-page":"599","article-title":"Horizontally Fused Training Array: An Effective Hardware Utilization Squeezer for Training Novel Deep Learning Models","volume":"3","author":"Wang Shang","year":"2021","unstructured":"Shang Wang , Peiming Yang , Yuxuan Zheng , Xin Li , and Gennady Pekhimenko . 2021 . Horizontally Fused Training Array: An Effective Hardware Utilization Squeezer for Training Novel Deep Learning Models . Proceedings of Machine Learning and Systems 3 (2021), 599 \u2013 623 . Shang Wang, Peiming Yang, Yuxuan Zheng, Xin Li, and Gennady Pekhimenko. 2021. Horizontally Fused Training Array: An Effective Hardware Utilization Squeezer for Training Novel Deep Learning Models. Proceedings of Machine Learning and Systems 3 (2021), 599\u2013623.","journal-title":"Proceedings of Machine Learning and Systems"},{"key":"e_1_3_2_1_35_1","volume-title":"Yolop: You only look once for panoptic driving perception. arXiv preprint arXiv:2108.11250","author":"Wu Dong","year":"2021","unstructured":"Dong Wu , Manwen Liao , Weitian Zhang , and Xinggang Wang . 2021 . Yolop: You only look once for panoptic driving perception. arXiv preprint arXiv:2108.11250 (2021). Dong Wu, Manwen Liao, Weitian Zhang, and Xinggang Wang. 2021. Yolop: You only look once for panoptic driving perception. arXiv preprint arXiv:2108.11250 (2021)."},{"key":"e_1_3_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1109\/SEC50012.2020.00041"},{"key":"e_1_3_2_1_37_1","volume-title":"Automated Runtime-Aware Scheduling for Multi-Tenant DNN Inference on GPU. In 2021 IEEE\/ACM International Conference On Computer Aided Design (ICCAD). IEEE, 1\u20139.","author":"Yu Fuxun","year":"2021","unstructured":"Fuxun Yu , Shawn Bray , Di Wang , Longfei Shangguan , Xulong Tang , Chenchen Liu , and Xiang Chen . 2021 . Automated Runtime-Aware Scheduling for Multi-Tenant DNN Inference on GPU. In 2021 IEEE\/ACM International Conference On Computer Aided Design (ICCAD). IEEE, 1\u20139. Fuxun Yu, Shawn Bray, Di Wang, Longfei Shangguan, Xulong Tang, Chenchen Liu, and Xiang Chen. 2021. Automated Runtime-Aware Scheduling for Multi-Tenant DNN Inference on GPU. In 2021 IEEE\/ACM International Conference On Computer Aided Design (ICCAD). IEEE, 1\u20139."},{"key":"e_1_3_2_1_38_1","volume-title":"A Survey of Multi-Tenant Deep Learning Inference on GPU. arXiv preprint arXiv:2203.09040","author":"Yu Fuxun","year":"2022","unstructured":"Fuxun Yu , Di Wang , Longfei Shangguan , Minjia Zhang , Chenchen Liu , and Xiang Chen . 2022. A Survey of Multi-Tenant Deep Learning Inference on GPU. arXiv preprint arXiv:2203.09040 ( 2022 ). Fuxun Yu, Di Wang, Longfei Shangguan, Minjia Zhang, Chenchen Liu, and Xiang Chen. 2022. A Survey of Multi-Tenant Deep Learning Inference on GPU. arXiv preprint arXiv:2203.09040 (2022)."},{"key":"e_1_3_2_1_39_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2020.3013005"},{"key":"e_1_3_2_1_40_1","first-page":"848","article-title":"DietCode: Automatic Optimization for Dynamic Tensor Programs","volume":"4","author":"Zheng Bojian","year":"2022","unstructured":"Bojian Zheng , Ziheng Jiang , Cody\u00a0Hao Yu , Haichen Shen , Joshua Fromm , Yizhi Liu , Yida Wang , Luis Ceze , Tianqi Chen , and Gennady Pekhimenko . 2022 . DietCode: Automatic Optimization for Dynamic Tensor Programs . Proceedings of Machine Learning and Systems 4 (2022), 848 \u2013 863 . Bojian Zheng, Ziheng Jiang, Cody\u00a0Hao Yu, Haichen Shen, Joshua Fromm, Yizhi Liu, Yida Wang, Luis Ceze, Tianqi Chen, and Gennady Pekhimenko. 2022. DietCode: Automatic Optimization for Dynamic Tensor Programs. Proceedings of Machine Learning and Systems 4 (2022), 848\u2013863.","journal-title":"Proceedings of Machine Learning and Systems"},{"key":"e_1_3_2_1_41_1","volume-title":"14th USENIX symposium on operating systems design and implementation (OSDI 20)","author":"Zheng Lianmin","year":"2020","unstructured":"Lianmin Zheng , Chengfan Jia , Minmin Sun , Zhao Wu , Cody\u00a0Hao Yu , Ameer Haj-Ali , Yida Wang , Jun Yang , Danyang Zhuo , Koushik Sen , 2020 . Ansor: Generating { High-Performance} Tensor Programs for Deep Learning . In 14th USENIX symposium on operating systems design and implementation (OSDI 20) . 863\u2013879. Lianmin Zheng, Chengfan Jia, Minmin Sun, Zhao Wu, Cody\u00a0Hao Yu, Ameer Haj-Ali, Yida Wang, Jun Yang, Danyang Zhuo, Koushik Sen, 2020. Ansor: Generating { High-Performance} Tensor Programs for Deep Learning. In 14th USENIX symposium on operating systems design and implementation (OSDI 20). 863\u2013879."},{"key":"e_1_3_2_1_42_1","volume-title":"Alpa: Automating Inter- and Intra-Operator Parallelism for Distributed Deep Learning. In 16th USENIX Symposium on Operating Systems Design and Implementation (OSDI 22)","author":"Zheng Lianmin","year":"2022","unstructured":"Lianmin Zheng , Zhuohan Li , Hao Zhang , Yonghao Zhuang , Zhifeng Chen , Yanping Huang , Yida Wang , Yuanzhong Xu , Danyang Zhuo , Eric\u00a0 P. Xing , Joseph\u00a0 E. Gonzalez , and Ion Stoica . 2022 . Alpa: Automating Inter- and Intra-Operator Parallelism for Distributed Deep Learning. In 16th USENIX Symposium on Operating Systems Design and Implementation (OSDI 22) . 559\u2013578. Lianmin Zheng, Zhuohan Li, Hao Zhang, Yonghao Zhuang, Zhifeng Chen, Yanping Huang, Yida Wang, Yuanzhong Xu, Danyang Zhuo, Eric\u00a0P. Xing, Joseph\u00a0E. Gonzalez, and Ion Stoica. 2022. Alpa: Automating Inter- and Intra-Operator Parallelism for Distributed Deep Learning. In 16th USENIX Symposium on Operating Systems Design and Implementation (OSDI 22). 559\u2013578."}],"event":{"name":"IPSN '23: The 22nd International Conference on Information Processing in Sensor Networks","location":"San Antonio TX USA","acronym":"IPSN '23","sponsor":["SIGBED ACM Special Interest Group on Embedded Systems"]},"container-title":["The 22nd International Conference on Information Processing in Sensor Networks"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3583120.3586953","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3583120.3586953","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T17:48:48Z","timestamp":1750182528000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3583120.3586953"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,5,9]]},"references-count":42,"alternative-id":["10.1145\/3583120.3586953","10.1145\/3583120"],"URL":"https:\/\/doi.org\/10.1145\/3583120.3586953","relation":{},"subject":[],"published":{"date-parts":[[2023,5,9]]},"assertion":[{"value":"2023-05-09","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}