{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T04:31:43Z","timestamp":1750221103767,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":54,"publisher":"ACM","license":[{"start":{"date-parts":[[2019,3,25]],"date-time":"2019-03-25T00:00:00Z","timestamp":1553472000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/100000001","name":"National Science Foundation","doi-asserted-by":"publisher","award":["CCF-1629559, CCF-1725663"],"award-info":[{"award-number":["CCF-1629559, CCF-1725663"]}],"id":[{"id":"10.13039\/100000001","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2019,3,25]]},"DOI":"10.1145\/3302424.3303954","type":"proceedings-article","created":{"date-parts":[[2019,3,22]],"date-time":"2019-03-22T13:10:03Z","timestamp":1553260203000},"page":"1-17","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":4,"title":["Automating Dependence-Aware Parallelization of Machine Learning Training on Distributed Shared Memory"],"prefix":"10.1145","author":[{"given":"Jinliang","family":"Wei","sequence":"first","affiliation":[{"name":"Carnegie Mellon University"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Garth A.","family":"Gibson","sequence":"additional","affiliation":[{"name":"Vector Institute, CMU and University of Toronto"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Phillip B.","family":"Gibbons","sequence":"additional","affiliation":[{"name":"Carnegie Mellon University"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Eric P.","family":"Xing","sequence":"additional","affiliation":[{"name":"Petuum Inc., Carnegie Mellon University"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2019,3,25]]},"reference":[{"key":"e_1_3_2_1_1_1","unstructured":"2009. Netflix Prize Data. https:\/\/www.kaggle.com\/netflix-inc\/netflix-prize-data\/.  2009. Netflix Prize Data. https:\/\/www.kaggle.com\/netflix-inc\/netflix-prize-data\/."},{"key":"e_1_3_2_1_2_1","unstructured":"2013. ClueWeb. https:\/\/lemurproject.org\/clueweb12\/.  2013. ClueWeb. https:\/\/lemurproject.org\/clueweb12\/."},{"key":"e_1_3_2_1_3_1","unstructured":"Last visited Dec 2018. Julia Micro-Benchmark. https:\/\/julialang.org\/benchmarks\/.  Last visited Dec 2018. Julia Micro-Benchmark. https:\/\/julialang.org\/benchmarks\/."},{"key":"e_1_3_2_1_4_1","unstructured":"Last visited Dec 2018. MATLAB Parallel For Loop. https:\/\/www.mathworks.com\/help\/matlab\/ref\/parfor.html.  Last visited Dec 2018. MATLAB Parallel For Loop. https:\/\/www.mathworks.com\/help\/matlab\/ref\/parfor.html."},{"key":"e_1_3_2_1_5_1","volume-title":"12th USENIX Symposium on Operating Systems Design and Implementation (OSDI 16)","author":"Abadi Martin","year":"2016","unstructured":"Martin Abadi , Paul Barham , Jianmin Chen , Zhifeng Chen , Andy Davis , Jeffrey Dean , Matthieu Devin , Sanjay Ghemawat , Geoffrey Irving , Michael Isard , Manjunath Kudlur , Josh Levenberg , Rajat Monga , Sherry Moore , Derek G. Murray , Benoit Steiner , Paul Tucker , Vijay Vasudevan , Pete Warden , Martin Wicke , Yuan Yu , and Xiaoqiang Zheng . 2016 . TensorFlow: A system for large-scale machine learning . In 12th USENIX Symposium on Operating Systems Design and Implementation (OSDI 16) . 265--283. https:\/\/www.usenix.org\/system\/files\/conference\/osdi16\/osdi16-abadi.pdf Martin Abadi, Paul Barham, Jianmin Chen, Zhifeng Chen, Andy Davis, Jeffrey Dean, Matthieu Devin, Sanjay Ghemawat, Geoffrey Irving, Michael Isard, Manjunath Kudlur, Josh Levenberg, Rajat Monga, Sherry Moore, Derek G. Murray, Benoit Steiner, Paul Tucker, Vijay Vasudevan, Pete Warden, Martin Wicke, Yuan Yu, and Xiaoqiang Zheng. 2016. TensorFlow: A system for large-scale machine learning. In 12th USENIX Symposium on Operating Systems Design and Implementation (OSDI 16). 265--283. https:\/\/www.usenix.org\/system\/files\/conference\/osdi16\/osdi16-abadi.pdf"},{"key":"e_1_3_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1145\/29873.29875"},{"key":"e_1_3_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1137\/141000671"},{"key":"e_1_3_2_1_8_1","volume-title":"Jordan","author":"Blei David M.","year":"2003","unstructured":"David M. Blei , Andrew Y. Ng , and Michael I . Jordan . 2003 . Latent Dirichlet Allocation. J. Mach. Learn. Res . 3 (March 2003), 993--1022. http:\/\/dl.acm.org\/citation.cfm?id=944919.944937 David M. Blei, Andrew Y. Ng, and Michael I. Jordan. 2003. Latent Dirichlet Allocation. J. Mach. Learn. Res. 3 (March 2003), 993--1022. http:\/\/dl.acm.org\/citation.cfm?id=944919.944937"},{"key":"e_1_3_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.5555\/2738600.2738630"},{"key":"e_1_3_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1145\/2741948.2741970"},{"volume-title":"2014 USENIX Annual Technical Conference (USENIX ATC 14)","author":"Cui Henggang","key":"e_1_3_2_1_11_1","unstructured":"Henggang Cui , James Cipar , Qirong Ho , Jin Kyu Kim , Seunghak Lee , Abhimanu Kumar , Jinliang Wei , Wei Dai , Gregory R. Ganger , Phillip B. Gibbons , Garth A. Gibson , and Eric P. Xing . 2014. Exploiting Bounded Staleness to Speed Up Big Data Analytics . In 2014 USENIX Annual Technical Conference (USENIX ATC 14) . USENIX Association, Philadelphia, PA, 37--48. Henggang Cui, James Cipar, Qirong Ho, Jin Kyu Kim, Seunghak Lee, Abhimanu Kumar, Jinliang Wei, Wei Dai, Gregory R. Ganger, Phillip B. Gibbons, Garth A. Gibson, and Eric P. Xing. 2014. Exploiting Bounded Staleness to Speed Up Big Data Analytics. In 2014 USENIX Annual Technical Conference (USENIX ATC 14). USENIX Association, Philadelphia, PA, 37--48."},{"key":"e_1_3_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1145\/2670979.2670984"},{"key":"e_1_3_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1109\/99.660313"},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1006\/jpdc.1995.1105"},{"key":"e_1_3_2_1_15_1","volume-title":"Adaptive Subgradient Methods for Online Learning and Stochastic Optimization. J. Mach. Learn. Res. 12 (July","author":"Duchi John","year":"2011","unstructured":"John Duchi , Elad Hazan , and Yoram Singer . 2011. Adaptive Subgradient Methods for Online Learning and Stochastic Optimization. J. Mach. Learn. Res. 12 (July 2011 ), 2121--2159. http:\/\/dl.acm.org\/citation.cfm?id=1953048.2021068 John Duchi, Elad Hazan, and Yoram Singer. 2011. Adaptive Subgradient Methods for Online Learning and Stochastic Optimization. J. Mach. Learn. Res. 12 (July 2011), 2121--2159. http:\/\/dl.acm.org\/citation.cfm?id=1953048.2021068"},{"key":"e_1_3_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1007\/BF01407835"},{"key":"e_1_3_2_1_17_1","article-title":"Some efficient solutions to the affine scheduling problem. Part II. Multidimensional time","volume":"21","author":"Feautrier Paul","year":"1992","unstructured":"Paul Feautrier . 1992 . Some efficient solutions to the affine scheduling problem. Part II. Multidimensional time . International Journal of Parallel Programming 21 , 6 (01 Dec 1992), 389--420. Paul Feautrier. 1992. Some efficient solutions to the affine scheduling problem. Part II. Multidimensional time. International Journal of Parallel Programming 21, 6 (01 Dec 1992), 389--420.","journal-title":"International Journal of Parallel Programming"},{"key":"e_1_3_2_1_18_1","volume-title":"In JMLR Workshop and Conference Proceedings.","author":"Yu Hsiang","year":"2011","unstructured":"Hsiang fu Yu , Hung yi Lo , Hsun ping Hsieh , Jing kai Lou , Todd G. Mckenzie , Jung wei Chou , Po han Chung , Chia hua Ho , Chun fu Chang , Jui yu Weng , En syu Yan , Che wei Chang , Tsung ting Kuo , Po Tzu Chang , Chieh Po , Chien yuan Wang , Yi hung Huang , Yu xun Ruan , Yu shi Lin , Shou de Lin , Hsuan tien Lin , and Chih jen Lin . 2011 . Feature engineering and classifier ensemble for KDD Cup 2010 . In In JMLR Workshop and Conference Proceedings. Hsiang fu Yu, Hung yi Lo, Hsun ping Hsieh, Jing kai Lou, Todd G. Mckenzie, Jung wei Chou, Po han Chung, Chia hua Ho, Chun fu Chang, Jui yu Weng, En syu Yan, Che wei Chang, Tsung ting Kuo, Po Tzu Chang, Chieh Po, Chien yuan Wang, Yi hung Huang, Yu xun Ruan, Yu shi Lin, Shou de Lin, Hsuan tien Lin, and Chih jen Lin. 2011. Feature engineering and classifier ensemble for KDD Cup 2010. In In JMLR Workshop and Conference Proceedings."},{"key":"e_1_3_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1145\/2020408.2020426"},{"volume-title":"Presented as part of the 10th USENIX Symposium on Operating Systems Design and Implementation (OSDI 12)","author":"Gonzalez Joseph E.","key":"e_1_3_2_1_20_1","unstructured":"Joseph E. Gonzalez , Yucheng Low , Haijie Gu , Danny Bickson , and Carlos Guestrin . 2012. PowerGraph: Distributed Graph-Parallel Computation on Natural Graphs . In Presented as part of the 10th USENIX Symposium on Operating Systems Design and Implementation (OSDI 12) . USENIX , Hollywood, CA , 17--30. Joseph E. Gonzalez, Yucheng Low, Haijie Gu, Danny Bickson, and Carlos Guestrin. 2012. PowerGraph: Distributed Graph-Parallel Computation on Natural Graphs. In Presented as part of the 10th USENIX Symposium on Operating Systems Design and Implementation (OSDI 12). USENIX, Hollywood, CA, 17--30."},{"key":"e_1_3_2_1_21_1","volume-title":"Deep Residual Learning for Image Recognition. CoRR abs\/1512.03385","author":"He Kaiming","year":"2015","unstructured":"Kaiming He , Xiangyu Zhang , Shaoqing Ren , and Jian Sun . 2015. Deep Residual Learning for Image Recognition. CoRR abs\/1512.03385 ( 2015 ). arXiv:1512.03385 http:\/\/arxiv.org\/abs\/1512.03385 Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2015. Deep Residual Learning for Image Recognition. CoRR abs\/1512.03385 (2015). arXiv:1512.03385 http:\/\/arxiv.org\/abs\/1512.03385"},{"key":"e_1_3_2_1_22_1","volume-title":"Phillip B. Gibbons, Garth A Gibson, Greg Ganger, and Eric P Xing.","author":"Ho Qirong","year":"2013","unstructured":"Qirong Ho , James Cipar , Henggang Cui , Seunghak Lee , Jin Kyu Kim , Phillip B. Gibbons, Garth A Gibson, Greg Ganger, and Eric P Xing. 2013 . More Effective Distributed ML via a Stale Synchronous Parallel Parameter Server. In Advances in Neural Information Processing Systems 26, C.J.C. Burges, L. Bottou, M. Welling, Z. Ghahramani, and K.Q. Weinberger (Eds.). Curran Associates, Inc ., 1223--1231. Qirong Ho, James Cipar, Henggang Cui, Seunghak Lee, Jin Kyu Kim, Phillip B. Gibbons, Garth A Gibson, Greg Ganger, and Eric P Xing. 2013. More Effective Distributed ML via a Stale Synchronous Parallel Parameter Server. In Advances in Neural Information Processing Systems 26, C.J.C. Burges, L. Bottou, M. Welling, Z. Ghahramani, and K.Q. Weinberger (Eds.). Curran Associates, Inc., 1223--1231."},{"key":"e_1_3_2_1_23_1","volume-title":"Advances in Neural Information Processing Systems 30: Annual Conference on Neural Information Processing Systems 2017","author":"Hoffer Elad","year":"2017","unstructured":"Elad Hoffer , Itay Hubara , and Daniel Soudry . 2017 . Train longer, generalize better: closing the generalization gap in large batch training of neural networks . In Advances in Neural Information Processing Systems 30: Annual Conference on Neural Information Processing Systems 2017 , 4--9 December 2017, Long Beach, CA, USA. 1729--1739. Elad Hoffer, Itay Hubara, and Daniel Soudry. 2017. Train longer, generalize better: closing the generalization gap in large batch training of neural networks. In Advances in Neural Information Processing Systems 30: Annual Conference on Neural Information Processing Systems 2017, 4--9 December 2017, Long Beach, CA, USA. 1729--1739."},{"key":"e_1_3_2_1_24_1","volume-title":"Allen","author":"Kennedy Ken","year":"2002","unstructured":"Ken Kennedy and John R . Allen . 2002 . Optimizing Compilers for Modern Architectures: A Dependence-based Approach. Morgan Kaufmann Publishers Inc ., San Francisco, CA, USA. Ken Kennedy and John R. Allen. 2002. Optimizing Compilers for Modern Architectures: A Dependence-based Approach. Morgan Kaufmann Publishers Inc., San Francisco, CA, USA."},{"key":"e_1_3_2_1_25_1","volume-title":"On Large-Batch Training for Deep Learning: Generalization Gap and Sharp Minima. CoRR abs\/1609.04836","author":"Keskar Nitish Shirish","year":"2016","unstructured":"Nitish Shirish Keskar , Dheevatsa Mudigere , Jorge Nocedal , Mikhail Smelyanskiy , and Ping Tak Peter Tang . 2016. On Large-Batch Training for Deep Learning: Generalization Gap and Sharp Minima. CoRR abs\/1609.04836 ( 2016 ). arXiv:1609.04836 http:\/\/arxiv.org\/abs\/1609.04836 Nitish Shirish Keskar, Dheevatsa Mudigere, Jorge Nocedal, Mikhail Smelyanskiy, and Ping Tak Peter Tang. 2016. On Large-Batch Training for Deep Learning: Generalization Gap and Sharp Minima. CoRR abs\/1609.04836 (2016). arXiv:1609.04836 http:\/\/arxiv.org\/abs\/1609.04836"},{"key":"e_1_3_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1145\/2901318.2901331"},{"key":"e_1_3_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/MC.2009.263"},{"key":"e_1_3_2_1_28_1","volume-title":"Proceedings of the 25th International Conference on Neural Information Processing Systems -","volume":"1","author":"Krizhevsky Alex","unstructured":"Alex Krizhevsky , Ilya Sutskever , and Geoffrey E. Hinton . 2012. ImageNet Classification with Deep Convolutional Neural Networks . In Proceedings of the 25th International Conference on Neural Information Processing Systems - Volume 1 (NIPS'12). Curran Associates Inc., USA, 1097--1105. http:\/\/dl.acm.org\/citation.cfm?id=2999134.2999257 Alex Krizhevsky, Ilya Sutskever, and Geoffrey E. Hinton. 2012. ImageNet Classification with Deep Convolutional Neural Networks. In Proceedings of the 25th International Conference on Neural Information Processing Systems - Volume 1 (NIPS'12). Curran Associates Inc., USA, 1097--1105. http:\/\/dl.acm.org\/citation.cfm?id=2999134.2999257"},{"key":"e_1_3_2_1_29_1","volume-title":"Scaling Distributed Machine Learning with the Parameter Server. In 11th USENIX Symposium on Operating Systems Design and Implementation (OSDI 14)","author":"Li Mu","year":"2014","unstructured":"Mu Li , David G. Andersen , Jun Woo Park , Alexander J. Smola , Amr Ahmed , Vanja Josifovski , James Long , Eugene J. Shekita , and Bor-Yiing Su . 2014 . Scaling Distributed Machine Learning with the Parameter Server. In 11th USENIX Symposium on Operating Systems Design and Implementation (OSDI 14) . USENIX Association, Broomfield, CO, 583--598. Mu Li, David G. Andersen, Jun Woo Park, Alexander J. Smola, Amr Ahmed, Vanja Josifovski, James Long, Eugene J. Shekita, and Bor-Yiing Su. 2014. Scaling Distributed Machine Learning with the Parameter Server. In 11th USENIX Symposium on Operating Systems Design and Implementation (OSDI 14). USENIX Association, Broomfield, CO, 583--598."},{"key":"e_1_3_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1016\/S0167-8191(98)00021-0"},{"key":"e_1_3_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.14778\/2212351.2212354"},{"key":"e_1_3_2_1_32_1","volume-title":"Hellerstein","author":"Low Yucheng","year":"2010","unstructured":"Yucheng Low , Joseph Gonzalez , Aapo Kyrola , Danny Bickson , Carlos Guestrin , and Joseph M . Hellerstein . 2010 . GraphLab: A New Framework For Parallel Machine Learning. In UAI. Yucheng Low, Joseph Gonzalez, Aapo Kyrola, Danny Bickson, Carlos Guestrin, and Joseph M. Hellerstein. 2010. GraphLab: A New Framework For Parallel Machine Learning. In UAI."},{"key":"e_1_3_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.1145\/113445.113447"},{"key":"e_1_3_2_1_34_1","volume-title":"Delay-Tolerant Algorithms for Asynchronous Distributed Online Learning. Advances in Neural Information Processing Systems (NIPS)","author":"McMahan H. Brendan","year":"2014","unstructured":"H. Brendan McMahan and Matthew Streeter . 2014. Delay-Tolerant Algorithms for Asynchronous Distributed Online Learning. Advances in Neural Information Processing Systems (NIPS) ( 2014 ). H. Brendan McMahan and Matthew Streeter. 2014. Delay-Tolerant Algorithms for Asynchronous Distributed Online Learning. Advances in Neural Information Processing Systems (NIPS) (2014)."},{"key":"e_1_3_2_1_35_1","volume-title":"Jordan","author":"Moritz Philipp","year":"2015","unstructured":"Philipp Moritz , Robert Nishihara , Ion Stoica , and Michael I . Jordan . 2015 . SparkNet: Training Deep Networks in Spark. CoRR abs\/1511.06051 (2015). arXiv:1511.06051 http:\/\/arxiv.org\/abs\/1511.06051 Philipp Moritz, Robert Nishihara, Ion Stoica, and Michael I. Jordan. 2015. SparkNet: Training Deep Networks in Spark. CoRR abs\/1511.06051 (2015). arXiv:1511.06051 http:\/\/arxiv.org\/abs\/1511.06051"},{"key":"e_1_3_2_1_36_1","volume-title":"Proceedings of the 8th USENIX Conference on Networked Systems Design and Implementation (NSDI'11)","author":"Murray Derek G.","year":"2011","unstructured":"Derek G. Murray , Malte Schwarzkopf , Christopher Smowton , Steven Smith , Anil Madhavapeddy , and Steven Hand . 2011 . CIEL: A Universal Execution Engine for Distributed Data-flow Computing . In Proceedings of the 8th USENIX Conference on Networked Systems Design and Implementation (NSDI'11) . USENIX Association, Berkeley, CA, USA, 113--126. http:\/\/dl.acm.org\/citation.cfm?id= 1972457.1972470 Derek G. Murray, Malte Schwarzkopf, Christopher Smowton, Steven Smith, Anil Madhavapeddy, and Steven Hand. 2011. CIEL: A Universal Execution Engine for Distributed Data-flow Computing. In Proceedings of the 8th USENIX Conference on Networked Systems Design and Implementation (NSDI'11). USENIX Association, Berkeley, CA, USA, 113--126. http:\/\/dl.acm.org\/citation.cfm?id=1972457.1972470"},{"key":"e_1_3_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.1145\/1454115.1454119"},{"key":"e_1_3_2_1_38_1","doi-asserted-by":"publisher","DOI":"10.1145\/1993498.1993501"},{"key":"e_1_3_2_1_39_1","volume-title":"Hogwild: A Lock-Free Approach to Parallelizing Stochastic Gradient Descent. In Advances in Neural Information Processing Systems 24","author":"Recht Benjamin","year":"2011","unstructured":"Benjamin Recht , Christopher Re , Stephen Wright , and Feng Niu . 2011 . Hogwild: A Lock-Free Approach to Parallelizing Stochastic Gradient Descent. In Advances in Neural Information Processing Systems 24 , J. Shawe-Taylor, R.S. Zemel, P.L. Bartlett, F. Pereira, and K.Q. Weinberger (Eds.). Curran Associates, Inc. , 693--701. Benjamin Recht, Christopher Re, Stephen Wright, and Feng Niu. 2011. Hogwild: A Lock-Free Approach to Parallelizing Stochastic Gradient Descent. In Advances in Neural Information Processing Systems 24, J. Shawe-Taylor, R.S. Zemel, P.L. Bartlett, F. Pereira, and K.Q. Weinberger (Eds.). Curran Associates, Inc., 693--701."},{"key":"e_1_3_2_1_40_1","doi-asserted-by":"publisher","DOI":"10.1145\/1183401.1183447"},{"key":"e_1_3_2_1_41_1","doi-asserted-by":"publisher","DOI":"10.1145\/1993316.1993518"},{"volume-title":"Advances in Neural Information Processing Systems 31. Curran Associates","author":"Shazeer Noam","key":"e_1_3_2_1_42_1","unstructured":"Noam Shazeer , Youlong Cheng , Niki Parmar , Dustin Tran , Ashish Vaswani , Penporn Koanantakool , Peter Hawkins , HyoukJoong Lee , Mingsheng Hong , Cliff Young , Ryan Sepassi , and Blake Hechtman . 2018. Mesh-TensorFlow: Deep Learning for Supercomputers . In Advances in Neural Information Processing Systems 31. Curran Associates , Inc ., 10435--10444. Noam Shazeer, Youlong Cheng, Niki Parmar, Dustin Tran, Ashish Vaswani, Penporn Koanantakool, Peter Hawkins, HyoukJoong Lee, Mingsheng Hong, Cliff Young, Ryan Sepassi, and Blake Hechtman. 2018. Mesh-TensorFlow: Deep Learning for Supercomputers. In Advances in Neural Information Processing Systems 31. Curran Associates, Inc., 10435--10444."},{"key":"e_1_3_2_1_43_1","doi-asserted-by":"publisher","DOI":"10.1145\/2025113.2025133"},{"key":"e_1_3_2_1_44_1","volume-title":"Proceedings of the 19th International Conference on Artificial Intelligence and Statistics, AISTATS 2016","author":"Sra Suvrit","year":"2016","unstructured":"Suvrit Sra , Adams Wei Yu , Mu Li , and Alexander J. Smola . 2016. AdaDelay: Delay Adaptive Distributed Stochastic Optimization . In Proceedings of the 19th International Conference on Artificial Intelligence and Statistics, AISTATS 2016 , Cadiz, Spain , May 9-11, 2016 . 957--965. http:\/\/jmlr.org\/proceedings\/papers\/v51\/sra16.html Suvrit Sra, Adams Wei Yu, Mu Li, and Alexander J. Smola. 2016. AdaDelay: Delay Adaptive Distributed Stochastic Optimization. In Proceedings of the 19th International Conference on Artificial Intelligence and Statistics, AISTATS 2016, Cadiz, Spain, May 9-11, 2016. 957--965. http:\/\/jmlr.org\/proceedings\/papers\/v51\/sra16.html"},{"key":"e_1_3_2_1_45_1","doi-asserted-by":"publisher","DOI":"10.1145\/2806777.2806778"},{"key":"e_1_3_2_1_46_1","doi-asserted-by":"publisher","DOI":"10.1109\/71.97902"},{"key":"e_1_3_2_1_47_1","unstructured":"Michael Wolfe. 1986. Advanced Loop Interchanging. In ICPP.  Michael Wolfe. 1986. Advanced Loop Interchanging. In ICPP."},{"key":"e_1_3_2_1_48_1","doi-asserted-by":"publisher","DOI":"10.1007\/BF01407876"},{"key":"e_1_3_2_1_49_1","volume-title":"14th USENIX Symposium on Networked Systems Design and Implementation (NSDI 17)","author":"Xiao Wencong","year":"2017","unstructured":"Wencong Xiao , Jilong Xue , Youshan Miao , Zhen Li , Cheng Chen , Ming Wu , Wei Li , and Lidong Zhou . 2017 . Tux2: Distributed Graph Computation for Machine Learning . In 14th USENIX Symposium on Networked Systems Design and Implementation (NSDI 17) . USENIX Association, Boston, MA, 669--682. https:\/\/www.usenix.org\/conference\/nsdi17\/technical-sessions\/presentation\/xiao Wencong Xiao, Jilong Xue, Youshan Miao, Zhen Li, Cheng Chen, Ming Wu, Wei Li, and Lidong Zhou. 2017. Tux2: Distributed Graph Computation for Machine Learning. In 14th USENIX Symposium on Networked Systems Design and Implementation (NSDI 17). USENIX Association, Boston, MA, 669--682. https:\/\/www.usenix.org\/conference\/nsdi17\/technical-sessions\/presentation\/xiao"},{"key":"e_1_3_2_1_50_1","doi-asserted-by":"publisher","DOI":"10.1145\/3190508.3190551"},{"key":"e_1_3_2_1_51_1","volume-title":"Proceedings of the 8th USENIX Conference on Operating Systems Design and Implementation (OSDI'08)","author":"Yu Yuan","year":"2008","unstructured":"Yuan Yu , Michael Isard , Dennis Fetterly , Mihai Budiu , \u00dalfar Erlingsson , Pradeep Kumar Gunda , and Jon Currey . 2008 . DryadLINQ: A System for General-purpose Distributed Data-parallel Computing Using a High-level Language . In Proceedings of the 8th USENIX Conference on Operating Systems Design and Implementation (OSDI'08) . USENIX Association, Berkeley, CA, USA, 1--14. http:\/\/dl.acm.org\/citation.cfm?id= 1855741.1855742 Yuan Yu, Michael Isard, Dennis Fetterly, Mihai Budiu, \u00dalfar Erlingsson, Pradeep Kumar Gunda, and Jon Currey. 2008. DryadLINQ: A System for General-purpose Distributed Data-parallel Computing Using a High-level Language. In Proceedings of the 8th USENIX Conference on Operating Systems Design and Implementation (OSDI'08). USENIX Association, Berkeley, CA, USA, 1--14. http:\/\/dl.acm.org\/citation.cfm?id=1855741.1855742"},{"volume-title":"Presented as part of the 9th USENIX Symposium on Networked Systems Design and Implementation (NSDI 12)","author":"Zaharia Matei","key":"e_1_3_2_1_52_1","unstructured":"Matei Zaharia , Mosharaf Chowdhury , Tathagata Das , Ankur Dave , Justin Ma , Murphy McCauly , Michael J. Franklin , Scott Shenker , and Ion Stoica . 2012. Resilient Distributed Datasets: A Fault-Tolerant Abstraction for In-Memory Cluster Computing . In Presented as part of the 9th USENIX Symposium on Networked Systems Design and Implementation (NSDI 12) . USENIX , San Jose, CA , 15--28. Matei Zaharia, Mosharaf Chowdhury, Tathagata Das, Ankur Dave, Justin Ma, Murphy McCauly, Michael J. Franklin, Scott Shenker, and Ion Stoica. 2012. Resilient Distributed Datasets: A Fault-Tolerant Abstraction for In-Memory Cluster Computing. In Presented as part of the 9th USENIX Symposium on Networked Systems Design and Implementation (NSDI 12). USENIX, San Jose, CA, 15--28."},{"key":"e_1_3_2_1_53_1","volume-title":"Exploring the Hidden Dimension in Graph Processing. In 12th USENIX Symposium on Operating Systems Design and Implementation (OSDI 16)","author":"Zhang Mingxing","year":"2016","unstructured":"Mingxing Zhang , Yongwei Wu , Kang Chen , Xuehai Qian , Xue Li , and Weimin Zheng . 2016 . Exploring the Hidden Dimension in Graph Processing. In 12th USENIX Symposium on Operating Systems Design and Implementation (OSDI 16) . USENIX Association, Savannah, GA, 285--300. https:\/\/www.usenix.org\/conference\/osdi16\/technical-sessions\/presentation\/zhang-mingxing Mingxing Zhang, Yongwei Wu, Kang Chen, Xuehai Qian, Xue Li, and Weimin Zheng. 2016. Exploring the Hidden Dimension in Graph Processing. In 12th USENIX Symposium on Operating Systems Design and Implementation (OSDI 16). USENIX Association, Savannah, GA, 285--300. https:\/\/www.usenix.org\/conference\/osdi16\/technical-sessions\/presentation\/zhang-mingxing"},{"key":"e_1_3_2_1_54_1","volume-title":"Gemini: A Computation-Centric Distributed Graph Processing System. In 12th USENIX Symposium on Operating Systems Design and Implementation (OSDI 16)","author":"Zhu Xiaowei","year":"2016","unstructured":"Xiaowei Zhu , Wenguang Chen , Weimin Zheng , and Xiaosong Ma . 2016 . Gemini: A Computation-Centric Distributed Graph Processing System. In 12th USENIX Symposium on Operating Systems Design and Implementation (OSDI 16) . USENIX Association, Savannah, GA, 301--316. https:\/\/www.usenix.org\/conference\/osdi16\/technical-sessions\/presentation\/zhu Xiaowei Zhu, Wenguang Chen, Weimin Zheng, and Xiaosong Ma. 2016. Gemini: A Computation-Centric Distributed Graph Processing System. In 12th USENIX Symposium on Operating Systems Design and Implementation (OSDI 16). USENIX Association, Savannah, GA, 301--316. https:\/\/www.usenix.org\/conference\/osdi16\/technical-sessions\/presentation\/zhu"}],"event":{"name":"EuroSys '19: Fourteenth EuroSys Conference 2019","sponsor":["SIGOPS ACM Special Interest Group on Operating Systems"],"location":"Dresden Germany","acronym":"EuroSys '19"},"container-title":["Proceedings of the Fourteenth EuroSys Conference 2019"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3302424.3303954","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3302424.3303954","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3302424.3303954","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T01:01:48Z","timestamp":1750208508000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3302424.3303954"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2019,3,25]]},"references-count":54,"alternative-id":["10.1145\/3302424.3303954","10.1145\/3302424"],"URL":"https:\/\/doi.org\/10.1145\/3302424.3303954","relation":{},"subject":[],"published":{"date-parts":[[2019,3,25]]},"assertion":[{"value":"2019-03-25","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}