{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,2]],"date-time":"2026-06-02T14:37:06Z","timestamp":1780411026589,"version":"3.54.1"},"reference-count":35,"publisher":"Association for Computing Machinery (ACM)","issue":"1","license":[{"start":{"date-parts":[[2019,7,25]],"date-time":"2019-07-25T00:00:00Z","timestamp":1564012800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["SIGOPS Oper. Syst. Rev."],"published-print":{"date-parts":[[2019,7,25]]},"abstract":"<jats:p>The computational demands of modern AI techniques are immense, and as the number of practical applications grows, there will be an increasing burden on shared computing infrastructure. We envision a forthcoming era of \"AI Systems\" research where reducing resource consumption, reasoning about transient resource availability, trading off resource consumption for accuracy, and managing contention on specialized hardware will become the community's main research focus. This paper overviews the history of AI systems research, a vision for the future, and the open challenges ahead.<\/jats:p>","DOI":"10.1145\/3352020.3352022","type":"journal-article","created":{"date-parts":[[2019,7,26]],"date-time":"2019-07-26T13:17:18Z","timestamp":1564147038000},"page":"1-6","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":26,"title":["Artificial Intelligence in Resource-Constrained and Shared Environments"],"prefix":"10.1145","volume":"53","author":[{"given":"Sanjay","family":"Krishnan","sequence":"first","affiliation":[{"name":"University of Chicago, Chicago, IL, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Aaron J.","family":"Elmore","sequence":"additional","affiliation":[{"name":"University of Chicago, Chicago, IL, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Michael","family":"Franklin","sequence":"additional","affiliation":[{"name":"University of Chicago, Chicago, IL, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"John","family":"Paparrizos","sequence":"additional","affiliation":[{"name":"University of Chicago, Chicago, IL, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Zechao","family":"Shang","sequence":"additional","affiliation":[{"name":"University of Chicago, Chicago, IL, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Adam","family":"Dziedzic","sequence":"additional","affiliation":[{"name":"University of Chicago, Chicago, IL, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Rui","family":"Liu","sequence":"additional","affiliation":[{"name":"University of Chicago, Chicago, IL, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2019,7,25]]},"reference":[{"key":"e_1_2_1_1_1","first-page":"265","volume-title":"12th {USENIX} Symposium on Operating Systems Design and Implementation ({OSDI} 16)","author":"Abadi M.","year":"2016","unstructured":"M. Abadi , P. Barham , J. Chen , Z. Chen , A. Davis , J. Dean , M. Devin , S. Ghemawat , G. Irving , M. Isard , et al. Tensorflow: A system for largescale machine learning . In 12th {USENIX} Symposium on Operating Systems Design and Implementation ({OSDI} 16) , pages 265 -- 283 , 2016 . M. Abadi, P. Barham, J. Chen, Z. Chen, A. Davis, J. Dean, M. Devin, S. Ghemawat, G. Irving, M. Isard, et al. Tensorflow: A system for largescale machine learning. In 12th {USENIX} Symposium on Operating Systems Design and Implementation ({OSDI} 16), pages 265--283, 2016."},{"key":"e_1_2_1_2_1","volume-title":"An analysis of deep neural network models for practical applications. arXiv preprint arXiv:1605.07678","author":"Canziani A.","year":"2016","unstructured":"A. Canziani , A. Paszke , and E. Culurciello . An analysis of deep neural network models for practical applications. arXiv preprint arXiv:1605.07678 , 2016 . A. Canziani, A. Paszke, and E. Culurciello. An analysis of deep neural network models for practical applications. arXiv preprint arXiv:1605.07678, 2016."},{"key":"e_1_2_1_3_1","first-page":"1","volume-title":"Tvm: end-to-end optimization stack for deep learning. arXiv preprint arXiv:1802.04799","author":"Chen T.","year":"2018","unstructured":"T. Chen , T. Moreau , Z. Jiang , H. Shen , E. Q. Yan , L. Wang , Y. Hu , L. Ceze , C. Guestrin , and A. Krishnamurthy . Tvm: end-to-end optimization stack for deep learning. arXiv preprint arXiv:1802.04799 , pages 1 -- 15 , 2018 . T. Chen, T. Moreau, Z. Jiang, H. Shen, E. Q. Yan, L.Wang, Y. Hu, L. Ceze, C. Guestrin, and A. Krishnamurthy. Tvm: end-to-end optimization stack for deep learning. arXiv preprint arXiv:1802.04799, pages 1--15, 2018."},{"key":"e_1_2_1_4_1","volume-title":"Training deep nets with sublinear memory cost. arXiv preprint arXiv:1604.06174","author":"Chen T.","year":"2016","unstructured":"T. Chen , B. Xu , C. Zhang , and C. Guestrin . Training deep nets with sublinear memory cost. arXiv preprint arXiv:1604.06174 , 2016 . T. Chen, B. Xu, C. Zhang, and C. Guestrin. Training deep nets with sublinear memory cost. arXiv preprint arXiv:1604.06174, 2016."},{"key":"e_1_2_1_5_1","volume-title":"cudnn: Efficient primitives for deep learning. arXiv preprint arXiv:1410.0759","author":"Chetlur S.","year":"2014","unstructured":"S. Chetlur , C. Woolley , P. Vandermersch , J. Cohen , J. Tran , B. Catanzaro , and E. Shelhamer . cudnn: Efficient primitives for deep learning. arXiv preprint arXiv:1410.0759 , 2014 . S. Chetlur, C.Woolley, P. Vandermersch, J. Cohen, J. Tran, B. Catanzaro, and E. Shelhamer. cudnn: Efficient primitives for deep learning. arXiv preprint arXiv:1410.0759, 2014."},{"key":"e_1_2_1_6_1","first-page":"613","volume-title":"14th {USENIX} Symposium on Networked Systems Design and Implementation ({NSDI} 17)","author":"Crankshaw D.","year":"2017","unstructured":"D. Crankshaw , X. Wang , G. Zhou , M. J. Franklin , J. E. Gonzalez , and I. Stoica . Clipper: A low-latency online prediction serving system . In 14th {USENIX} Symposium on Networked Systems Design and Implementation ({NSDI} 17) , pages 613 -- 627 , 2017 . D. Crankshaw, X. Wang, G. Zhou, M. J. Franklin, J. E. Gonzalez, and I. Stoica. Clipper: A low-latency online prediction serving system. In 14th {USENIX} Symposium on Networked Systems Design and Implementation ({NSDI} 17), pages 613--627, 2017."},{"key":"e_1_2_1_7_1","volume-title":"GitHub repository","author":"Dhariwal P.","year":"2017","unstructured":"P. Dhariwal , C. Hesse , O. Klimov , A. Nichol , M. Plappert , A. Radford , J. Schulman , S. Sidor , Y. Wu , and P. Zhokhov . Openai baselines. GitHub , GitHub repository , 2017 . P. Dhariwal, C. Hesse, O. Klimov, A. Nichol, M. Plappert, A. Radford, J. Schulman, S. Sidor, Y.Wu, and P. Zhokhov. Openai baselines. GitHub, GitHub repository, 2017."},{"key":"e_1_2_1_8_1","volume-title":"ICML","author":"A.","year":"2019","unstructured":"A. Dziedzic*, J. Paparrizos*, S. Krishnan , A. Elmore , and M. Franklin . Band-limited training and inference for convolutional neural networks . ICML , 2019 . A. Dziedzic*, J. Paparrizos*, S. Krishnan, A. Elmore, and M. Franklin. Band-limited training and inference for convolutional neural networks. ICML, 2019."},{"key":"e_1_2_1_9_1","volume-title":"Proc. of NetDB Workshop","author":"Elmore A. J.","year":"2011","unstructured":"A. J. Elmore , S. Das , D. Agrawal , and A. El Abbadi . Towards an elastic and autonomic multitenant database . In Proc. of NetDB Workshop , 2011 . A. J. Elmore, S. Das, D. Agrawal, and A. El Abbadi. Towards an elastic and autonomic multitenant database. In Proc. of NetDB Workshop, 2011."},{"key":"e_1_2_1_10_1","first-page":"2214","volume-title":"Advances in neural information processing systems","author":"Gomez A. N.","year":"2017","unstructured":"A. N. Gomez , M. Ren , R. Urtasun , and R. B. Grosse . The reversible residual network: Backpropagation without storing activations . In Advances in neural information processing systems , pages 2214 -- 2224 , 2017 . A. N. Gomez, M. Ren, R. Urtasun, and R. B. Grosse. The reversible residual network: Backpropagation without storing activations. In Advances in neural information processing systems, pages 2214--2224, 2017."},{"key":"e_1_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01234-2_48"},{"key":"e_1_2_1_12_1","volume-title":"Distilling the knowledge in a neural network. arXiv preprint arXiv:1503.02531","author":"Hinton G.","year":"2015","unstructured":"G. Hinton , O. Vinyals , and J. Dean . Distilling the knowledge in a neural network. arXiv preprint arXiv:1503.02531 , 2015 . G. Hinton, O. Vinyals, and J. Dean. Distilling the knowledge in a neural network. arXiv preprint arXiv:1503.02531, 2015."},{"key":"e_1_2_1_13_1","volume-title":"Mobilenets: Efficient convolutional neural networks for mobile vision applications. arXiv preprint arXiv:1704.04861","author":"Howard A. G.","year":"2017","unstructured":"A. G. Howard , M. Zhu , B. Chen , D. Kalenichenko , W. Wang , T. Weyand , M. Andreetto , and H. Adam . Mobilenets: Efficient convolutional neural networks for mobile vision applications. arXiv preprint arXiv:1704.04861 , 2017 . A. G. Howard, M. Zhu, B. Chen, D. Kalenichenko,W.Wang, T.Weyand, M. Andreetto, and H. Adam. Mobilenets: Efficient convolutional neural networks for mobile vision applications. arXiv preprint arXiv:1704.04861, 2017."},{"key":"e_1_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.284"},{"key":"e_1_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1109\/MM.2018.032271057"},{"key":"e_1_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1145\/1296907.1296909"},{"key":"e_1_2_1_17_1","first-page":"1097","volume-title":"Advances in neural information processing systems","author":"Krizhevsky A.","year":"2012","unstructured":"A. Krizhevsky , I. Sutskever , and G. E. Hinton . Imagenet classification with deep convolutional neural networks . In Advances in neural information processing systems , pages 1097 -- 1105 , 2012 . A. Krizhevsky, I. Sutskever, and G. E. Hinton. Imagenet classification with deep convolutional neural networks. In Advances in neural information processing systems, pages 1097--1105, 2012."},{"key":"e_1_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1145\/2723372.2742788"},{"key":"e_1_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1145\/2320765.2320789"},{"key":"e_1_2_1_20_1","first-page":"561","volume-title":"13th {USENIX} Symposium on Operating Systems Design and Implementation ({OSDI} 18)","author":"Moritz P.","year":"2018","unstructured":"P. Moritz , R. Nishihara , S. Wang , A. Tumanov , R. Liaw , E. Liang , M. Elibol , Z. Yang , W. Paul , M. I. Jordan , et al. Ray: A distributed framework for emerging {AI} applications . In 13th {USENIX} Symposium on Operating Systems Design and Implementation ({OSDI} 18) , pages 561 -- 577 , 2018 . P. Moritz, R. Nishihara, S.Wang, A. Tumanov, R. Liaw, E. Liang, M. Elibol, Z. Yang, W. Paul, M. I. Jordan, et al. Ray: A distributed framework for emerging {AI} applications. In 13th {USENIX} Symposium on Operating Systems Design and Implementation ({OSDI} 18), pages 561--577, 2018."},{"key":"e_1_2_1_21_1","volume-title":"Tensorflow-serving: Flexible, high-performance ml serving. arXiv preprint arXiv:1712.06139","author":"Olston C.","year":"2017","unstructured":"C. Olston , N. Fiedel , K. Gorovoy , J. Harmsen , L. Lao , F. Li , V. Rajashekhar , S. Ramesh , and J. Soyke . Tensorflow-serving: Flexible, high-performance ml serving. arXiv preprint arXiv:1712.06139 , 2017 . C. Olston, N. Fiedel, K. Gorovoy, J. Harmsen, L. Lao, F. Li, V. Rajashekhar, S. Ramesh, and J. Soyke. Tensorflow-serving: Flexible, high-performance ml serving. arXiv preprint arXiv:1712.06139, 2017."},{"key":"e_1_2_1_22_1","volume-title":"Pytorch: Tensors and dynamic neural networks in python with strong gpu acceleration. PyTorch: Tensors and dynamic neural networks in Python with strong GPU acceleration, 6","author":"Paszke A.","year":"2017","unstructured":"A. Paszke , S. Gross , S. Chintala , and G. Chanan . Pytorch: Tensors and dynamic neural networks in python with strong gpu acceleration. PyTorch: Tensors and dynamic neural networks in Python with strong GPU acceleration, 6 , 2017 . A. Paszke, S. Gross, S. Chintala, and G. Chanan. Pytorch: Tensors and dynamic neural networks in python with strong gpu acceleration. PyTorch: Tensors and dynamic neural networks in Python with strong GPU acceleration, 6, 2017."},{"key":"e_1_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1145\/2814576.2814808"},{"key":"e_1_2_1_24_1","unstructured":"A. Radford J. Wu R. Child D. Luan D. Amodei and I. Sutskever. Language models are unsupervised multitask learners. A. Radford J. Wu R. Child D. Luan D. Amodei and I. Sutskever. Language models are unsupervised multitask learners."},{"key":"e_1_2_1_25_1","volume-title":"et al. Sysml: The new frontier of machine learning systems. arXiv preprint arXiv:1904.03257","author":"Ratner A.","year":"2019","unstructured":"A. Ratner , D. Alistarh , G. Alonso , P. Bailis , S. Bird , N. Carlini , B. Catanzaro , E. Chung , B. Dally , J. Dean , et al. Sysml: The new frontier of machine learning systems. arXiv preprint arXiv:1904.03257 , 2019 . A. Ratner, D. Alistarh, G. Alonso, P. Bailis, S. Bird, N. Carlini, B. Catanzaro, E. Chung, B. Dally, J. Dean, et al. Sysml: The new frontier of machine learning systems. arXiv preprint arXiv:1904.03257, 2019."},{"key":"e_1_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1002\/j.1538-7305.1978.tb02136.x"},{"key":"e_1_2_1_27_1","volume-title":"Measuring the effects of data parallelism on neural network training. CoRR, abs\/1811.03600","author":"Shallue C. J.","year":"2018","unstructured":"C. J. Shallue , J. Lee , J. M. Antognini , J. Sohl-Dickstein , R. Frostig , and G. E. Dahl . Measuring the effects of data parallelism on neural network training. CoRR, abs\/1811.03600 , 2018 . C. J. Shallue, J. Lee, J. M. Antognini, J. Sohl-Dickstein, R. Frostig, and G. E. Dahl. Measuring the effects of data parallelism on neural network training. CoRR, abs\/1811.03600, 2018."},{"key":"e_1_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1038\/nature24270"},{"key":"e_1_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1145\/1272998.1273025"},{"key":"e_1_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1109\/CLOUD.2009.78"},{"key":"e_1_2_1_31_1","volume-title":"Pearson","author":"Tanenbaum A. S.","year":"2015","unstructured":"A. S. Tanenbaum and H. Bos . Modern operating systems . Pearson , 2015 . A. S. Tanenbaum and H. Bos. Modern operating systems. Pearson, 2015."},{"key":"e_1_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.5555\/1315451.1315479"},{"key":"e_1_2_1_33_1","volume-title":"et al. Theano: A python framework for fast computation of mathematical expressions. arXiv preprint arXiv:1605.02688","author":"Team T. T. D.","year":"2016","unstructured":"T. T. D. Team , R. Al-Rfou , G. Alain , A. Almahairi , C. Angermueller , D. Bahdanau , N. Ballas , F. Bastien , J. Bayer , A. Belikov , et al. Theano: A python framework for fast computation of mathematical expressions. arXiv preprint arXiv:1605.02688 , 2016 . T. T. D. Team, R. Al-Rfou, G. Alain, A. Almahairi, C. Angermueller, D. Bahdanau, N. Ballas, F. Bastien, J. Bayer, A. Belikov, et al. Theano: A python framework for fast computation of mathematical expressions. arXiv preprint arXiv:1605.02688, 2016."},{"key":"e_1_2_1_34_1","doi-asserted-by":"crossref","first-page":"167","DOI":"10.1007\/978-3-319-06486-4_7","volume-title":"High-Performance Computing on the Intel\u00ae Xeon Phi","author":"Wang E.","year":"2014","unstructured":"E. Wang , Q. Zhang , B. Shen , G. Zhang , X. Lu , Q. Wu , and Y. Wang . Intel math kernel library . In High-Performance Computing on the Intel\u00ae Xeon Phi , pages 167 -- 188 . Springer , 2014 . E. Wang, Q. Zhang, B. Shen, G. Zhang, X. Lu, Q. Wu, and Y. Wang. Intel math kernel library. In High-Performance Computing on the Intel\u00ae Xeon Phi, pages 167--188. Springer, 2014."},{"key":"e_1_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1145\/2966986.2967011"}],"container-title":["ACM SIGOPS Operating Systems Review"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3352020.3352022","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3352020.3352022","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T00:26:15Z","timestamp":1750206375000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3352020.3352022"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2019,7,25]]},"references-count":35,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2019,7,25]]}},"alternative-id":["10.1145\/3352020.3352022"],"URL":"https:\/\/doi.org\/10.1145\/3352020.3352022","relation":{},"ISSN":["0163-5980"],"issn-type":[{"value":"0163-5980","type":"print"}],"subject":[],"published":{"date-parts":[[2019,7,25]]},"assertion":[{"value":"2019-07-25","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}