{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,6]],"date-time":"2026-05-06T10:05:54Z","timestamp":1778061954836,"version":"3.51.4"},"reference-count":126,"publisher":"Association for Computing Machinery (ACM)","issue":"1","license":[{"start":{"date-parts":[[2021,1,2]],"date-time":"2021-01-02T00:00:00Z","timestamp":1609545600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/501100012166","name":"National Key R&D Program of China","doi-asserted-by":"crossref","award":["2019YFB2102200"],"award-info":[{"award-number":["2019YFB2102200"]}],"id":[{"id":"10.13039\/501100012166","id-type":"DOI","asserted-by":"crossref"}]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"crossref","award":["61972252, 61972254, 61672348, 61672353"],"award-info":[{"award-number":["61972252, 61972254, 61672348, 61672353"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"crossref"}]},{"name":"Joint Scientific Research Foundation of the State Education Ministry","award":["6141A02033702"],"award-info":[{"award-number":["6141A02033702"]}]},{"name":"Alibaba Innovation Research Program"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Comput. Surv."],"published-print":{"date-parts":[[2022,1,31]]},"abstract":"<jats:p>In recent years, mobile devices have gained increasing development with stronger computation capability and larger storage space. Some of the computation-intensive machine learning tasks can now be run on mobile devices. To exploit the resources available on mobile devices and preserve personal privacy, the concept of client-based machine learning has been proposed. It leverages the users\u2019 local hardware and local data to solve machine learning sub-problems on mobile devices and only uploads computation results rather than the original data for the optimization of the global model. Such an architecture can not only relieve computation and storage burdens on servers but also protect the users\u2019 sensitive information. Another benefit is the bandwidth reduction because various kinds of local data can be involved in the training process without being uploaded. In this article, we provide a literature review on the progressive development of machine learning from server based to client based. We revisit a number of widely used server-based and client-based machine learning methods and applications. We also extensively discuss the challenges and future directions in this area. We believe that this survey will give a clear overview of client-based machine learning and provide guidelines on applying client-based machine learning to practice.<\/jats:p>","DOI":"10.1145\/3424660","type":"journal-article","created":{"date-parts":[[2021,1,2]],"date-time":"2021-01-02T17:08:21Z","timestamp":1609607301000},"page":"1-36","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":24,"title":["From Server-Based to Client-Based Machine Learning"],"prefix":"10.1145","volume":"54","author":[{"ORCID":"https:\/\/orcid.org\/0000-0003-2895-444X","authenticated-orcid":false,"given":"Renjie","family":"Gu","sequence":"first","affiliation":[{"name":"Shanghai Jiao Tong University, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Chaoyue","family":"Niu","sequence":"additional","affiliation":[{"name":"Shanghai Jiao Tong University, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Fan","family":"Wu","sequence":"additional","affiliation":[{"name":"Shanghai Jiao Tong University, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Guihai","family":"Chen","sequence":"additional","affiliation":[{"name":"Shanghai Jiao Tong University, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Chun","family":"Hu","sequence":"additional","affiliation":[{"name":"Alibaba Group, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Chengfei","family":"Lyu","sequence":"additional","affiliation":[{"name":"Alibaba Group, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Zhihua","family":"Wu","sequence":"additional","affiliation":[{"name":"Alibaba Group, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2021,1,2]]},"reference":[{"key":"e_1_2_1_1_1","volume-title":"et\u00a0al","author":"Abadi Mart\u00edn","year":"2015","unstructured":"Mart\u00edn Abadi , Ashish Agarwal , Paul Barham , et\u00a0al . 2015 . TensorFlow: Large- Scale Machine Learning on Heterogeneous Systems. Retrieved from https:\/\/www.tensorflow.org\/. Software available from tensorflow.org. Mart\u00edn Abadi, Ashish Agarwal, Paul Barham, et\u00a0al. 2015. TensorFlow: Large-Scale Machine Learning on Heterogeneous Systems. Retrieved from https:\/\/www.tensorflow.org\/. Software available from tensorflow.org."},{"key":"e_1_2_1_2_1","volume-title":"Can we use split learning on 1D CNN models for privacy preserving training? arXiv preprint arXiv:2003.12365","author":"Abuadbba Sharif","year":"2020","unstructured":"Sharif Abuadbba , Kyuyeon Kim , Minki Kim , Chandra Thapa , Seyit A. Camtepe , Yansong Gao , Hyoungshick Kim , and Surya Nepal . 2020. Can we use split learning on 1D CNN models for privacy preserving training? arXiv preprint arXiv:2003.12365 ( 2020 ). Sharif Abuadbba, Kyuyeon Kim, Minki Kim, Chandra Thapa, Seyit A. Camtepe, Yansong Gao, Hyoungshick Kim, and Surya Nepal. 2020. Can we use split learning on 1D CNN models for privacy preserving training? arXiv preprint arXiv:2003.12365 (2020)."},{"key":"e_1_2_1_3_1","volume-title":"Felix Xinnan X. Yu, Sanjiv Kumar, and Brendan McMahan.","author":"Agarwal Naman","year":"2018","unstructured":"Naman Agarwal , Ananda Theertha Suresh , Felix Xinnan X. Yu, Sanjiv Kumar, and Brendan McMahan. 2018 . cpSGD: Communication-efficient and differentially-private distributed SGD. In Advances in Neural Information Processing Systems 31 (NeurIPS'18). Curran Associates, Inc ., 7564--7575. Naman Agarwal, Ananda Theertha Suresh, Felix Xinnan X. Yu, Sanjiv Kumar, and Brendan McMahan. 2018. cpSGD: Communication-efficient and differentially-private distributed SGD. In Advances in Neural Information Processing Systems 31 (NeurIPS'18). Curran Associates, Inc., 7564--7575."},{"key":"e_1_2_1_4_1","volume-title":"Proceedings of the 5th ACM International Conference on Web Search and Data Mining (WSDM'12)","author":"Ahmed Amr","unstructured":"Amr Ahmed , Moahmed Aly , Joseph Gonzalez , Shravan Narayanamurthy , and Alexander J. Smola . 2012. Scalable inference in latent variable models . In Proceedings of the 5th ACM International Conference on Web Search and Data Mining (WSDM'12) . 123--132. Amr Ahmed, Moahmed Aly, Joseph Gonzalez, Shravan Narayanamurthy, and Alexander J. Smola. 2012. Scalable inference in latent variable models. In Proceedings of the 5th ACM International Conference on Web Search and Data Mining (WSDM'12). 123--132."},{"key":"e_1_2_1_5_1","volume-title":"Proceedings of the 36th International Conference on Machine Learning (ICML'19)","volume":"97","author":"Amin Kareem","year":"2019","unstructured":"Kareem Amin , Alex Kulesza , Andres Munoz , and Sergei Vassilvtiskii . 2019 . Bounding user contributions: A bias-variance trade-off in differential privacy . In Proceedings of the 36th International Conference on Machine Learning (ICML'19) , Vol. 97 . 263--271. Kareem Amin, Alex Kulesza, Andres Munoz, and Sergei Vassilvtiskii. 2019. Bounding user contributions: A bias-variance trade-off in differential privacy. In Proceedings of the 36th International Conference on Machine Learning (ICML'19), Vol. 97. 263--271."},{"key":"e_1_2_1_6_1","volume-title":"Hinton","author":"Anil Rohan","year":"2018","unstructured":"Rohan Anil , Gabriel Pereyra , Alexandre Passos , Robert Ormandi , George E. Dahl , and Geoffrey E . Hinton . 2018 . Large scale distributed neural network training through online distillation. arXiv preprint arXiv:1804.03235 (2018). Rohan Anil, Gabriel Pereyra, Alexandre Passos, Robert Ormandi, George E. Dahl, and Geoffrey E. Hinton. 2018. Large scale distributed neural network training through online distillation. arXiv preprint arXiv:1804.03235 (2018)."},{"key":"e_1_2_1_7_1","volume-title":"Retrieved","year":"2017","unstructured":"Apple. 2017 . Core ML . Retrieved January 20, 2020, from https:\/\/developer.apple.com\/documentation\/coreml. Apple. 2017. Core ML. Retrieved January 20, 2020, from https:\/\/developer.apple.com\/documentation\/coreml."},{"key":"e_1_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1145\/2814566"},{"key":"e_1_2_1_9_1","volume-title":"How to backdoor federated learning. arXiv preprint arXiv:1807.00459","author":"Bagdasaryan Eugene","year":"2018","unstructured":"Eugene Bagdasaryan , Andreas Veit , Yiqing Hua , Deborah Estrin , and Vitaly Shmatikov . 2018. How to backdoor federated learning. arXiv preprint arXiv:1807.00459 ( 2018 ). Eugene Bagdasaryan, Andreas Veit, Yiqing Hua, Deborah Estrin, and Vitaly Shmatikov. 2018. How to backdoor federated learning. arXiv preprint arXiv:1807.00459 (2018)."},{"key":"e_1_2_1_10_1","volume-title":"Retrieved","year":"2019","unstructured":"Baidu. 2019 . Paddle Lite . Retrieved January 20, 2020, from https:\/\/github.com\/PaddlePaddle\/Paddle-Lite. Baidu. 2019. Paddle Lite. Retrieved January 20, 2020, from https:\/\/github.com\/PaddlePaddle\/Paddle-Lite."},{"key":"e_1_2_1_11_1","volume-title":"Pattern Recognition and Machine Learning","author":"Bishop Christopher M.","unstructured":"Christopher M. Bishop . 2006. Pattern Recognition and Machine Learning . Springer . Christopher M. Bishop. 2006. Pattern Recognition and Machine Learning. Springer."},{"key":"e_1_2_1_12_1","volume-title":"David Petrou, Daniel Ramage, and Jason Roselander.","author":"Bonawitz Keith","year":"2019","unstructured":"Keith Bonawitz , Hubert Eichner , Wolfgang Grieskamp , Dzmitry Huba , Alex Ingerman , Vladimir Ivanov , Chlo\u00e9 Kiddon , Jakub Konecn\u00fd , Stefano Mazzocchi , H. Brendan McMahan , Timon Van Overveldt , David Petrou, Daniel Ramage, and Jason Roselander. 2019 . Towards federated learning at scale: System de sign. arXiv preprint arXiv:1902.01046 (2019). Keith Bonawitz, Hubert Eichner, Wolfgang Grieskamp, Dzmitry Huba, Alex Ingerman, Vladimir Ivanov, Chlo\u00e9 Kiddon, Jakub Konecn\u00fd, Stefano Mazzocchi, H. Brendan McMahan, Timon Van Overveldt, David Petrou, Daniel Ramage, and Jason Roselander. 2019. Towards federated learning at scale: System design. arXiv preprint arXiv:1902.01046 (2019)."},{"key":"e_1_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1145\/3133956.3133982"},{"key":"e_1_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1137\/16M1080173"},{"key":"e_1_2_1_15_1","volume-title":"Distributed optimization and statistical learning via the alternating direction method of multipliers. Foundations and Trends\u00ae in Machine learning 3, 1","author":"Boyd Stephen","year":"2011","unstructured":"Stephen Boyd , Neal Parikh , Eric Chu , Borja Peleato , and Jonathan Eckstein . 2011. Distributed optimization and statistical learning via the alternating direction method of multipliers. Foundations and Trends\u00ae in Machine learning 3, 1 ( 2011 ), 1--122. Stephen Boyd, Neal Parikh, Eric Chu, Borja Peleato, and Jonathan Eckstein. 2011. Distributed optimization and statistical learning via the alternating direction method of multipliers. Foundations and Trends\u00ae in Machine learning 3, 1 (2011), 1--122."},{"key":"e_1_2_1_16_1","volume-title":"Expanding the reach of federated learning by reducing client resource requirements. arXiv preprint arXiv:1812.07210","author":"Caldas Sebastian","year":"2018","unstructured":"Sebastian Caldas , Jakub Kone\u010dny , H. Brendan McMahan , and Ameet Talwalkar . 2018. Expanding the reach of federated learning by reducing client resource requirements. arXiv preprint arXiv:1812.07210 ( 2018 ). Sebastian Caldas, Jakub Kone\u010dny, H. Brendan McMahan, and Ameet Talwalkar. 2018. Expanding the reach of federated learning by reducing client resource requirements. arXiv preprint arXiv:1812.07210 (2018)."},{"key":"e_1_2_1_17_1","volume-title":"Leaf: A benchmark for federated settings. arXiv preprint arXiv:1812.01097","author":"Caldas Sebastian","year":"2018","unstructured":"Sebastian Caldas , Peter Wu , Tian Li , Jakub Kone\u010dn\u1ef3 , H. Brendan McMahan , Virginia Smith , and Ameet Talwalkar . 2018 . Leaf: A benchmark for federated settings. arXiv preprint arXiv:1812.01097 (2018). Sebastian Caldas, Peter Wu, Tian Li, Jakub Kone\u010dn\u1ef3, H. Brendan McMahan, Virginia Smith, and Ameet Talwalkar. 2018. Leaf: A benchmark for federated settings. arXiv preprint arXiv:1812.01097 (2018)."},{"key":"e_1_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1145\/3307334.3326071"},{"key":"e_1_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.21437\/Interspeech.2017-1798"},{"key":"e_1_2_1_20_1","volume-title":"Proceedings of the 26th Annual International Conference on Mobile Computing and Networking (MobiCom'20)","author":"Chaoyue Niu","year":"2020","unstructured":"Niu Chaoyue , Wu Fan , Tang Shaojie , Hua Lifeng , Jia Rongfei , Lv Chengfei , Wu Zhihua , and Chen Guihai . 2020 . Billion-scale federated learning on mobile clients: A submodel design with tunable privacy . In Proceedings of the 26th Annual International Conference on Mobile Computing and Networking (MobiCom'20) . 405--418. Niu Chaoyue, Wu Fan, Tang Shaojie, Hua Lifeng, Jia Rongfei, Lv Chengfei, Wu Zhihua, and Chen Guihai. 2020. Billion-scale federated learning on mobile clients: A submodel design with tunable privacy. In Proceedings of the 26th Annual International Conference on Mobile Computing and Networking (MobiCom'20). 405--418."},{"key":"e_1_2_1_21_1","volume-title":"Federated meta-learning for recommendation. arXiv preprint arXiv:1802.07876","author":"Chen Fei","year":"2018","unstructured":"Fei Chen , Zhenhua Dong , Zhenguo Li , and Xiuqiang He. 2018. Federated meta-learning for recommendation. arXiv preprint arXiv:1802.07876 ( 2018 ). Fei Chen, Zhenhua Dong, Zhenguo Li, and Xiuqiang He. 2018. Federated meta-learning for recommendation. arXiv preprint arXiv:1802.07876 (2018)."},{"key":"e_1_2_1_22_1","volume-title":"Federated learning of out-of-vocabulary words. arXiv preprint arXiv:1903.10635","author":"Chen Mingqing","year":"2019","unstructured":"Mingqing Chen , Rajiv Mathews , Tom Ouyang , and Fran\u00e7oise Beaufays . 2019. Federated learning of out-of-vocabulary words. arXiv preprint arXiv:1903.10635 ( 2019 ). Mingqing Chen, Rajiv Mathews, Tom Ouyang, and Fran\u00e7oise Beaufays. 2019. Federated learning of out-of-vocabulary words. arXiv preprint arXiv:1903.10635 (2019)."},{"key":"e_1_2_1_23_1","volume-title":"SecureBoost: A lossless federated learning framework. arXiv preprint arXiv:1901.08755","author":"Cheng Kewei","year":"2019","unstructured":"Kewei Cheng , Tao Fan , Yilun Jin , Yang Liu , Tianjian Chen , and Qiang Yang . 2019. SecureBoost: A lossless federated learning framework. arXiv preprint arXiv:1901.08755 ( 2019 ). Kewei Cheng, Tao Fan, Yilun Jin, Yang Liu, Tianjian Chen, and Qiang Yang. 2019. SecureBoost: A lossless federated learning framework. arXiv preprint arXiv:1901.08755 (2019)."},{"key":"e_1_2_1_24_1","volume-title":"Ng","author":"Dean Jeffrey","year":"2012","unstructured":"Jeffrey Dean , Greg Corrado , Rajat Monga , Kai Chen , Matthieu Devin , Mark Mao , Marc'aurelio Ranzato , Andrew Senior , Paul Tucker , Ke Yang , Quoc V. Le , and Andrew Y . Ng . 2012 . Large scale distributed deep networks. In Advances in Neural Information Processing Systems 25 (NeurIPS'12). Curran Associates, Inc ., 1223--1231. Jeffrey Dean, Greg Corrado, Rajat Monga, Kai Chen, Matthieu Devin, Mark Mao, Marc'aurelio Ranzato, Andrew Senior, Paul Tucker, Ke Yang, Quoc V. Le, and Andrew Y. Ng. 2012. Large scale distributed deep networks. In Advances in Neural Information Processing Systems 25 (NeurIPS'12). Curran Associates, Inc., 1223--1231."},{"key":"e_1_2_1_25_1","unstructured":"Dheeru Dua and Casey Graff. 2017. UCI Machine Learning Repository. Retrieved from http:\/\/archive.ics.uci.edu\/ml.  Dheeru Dua and Casey Graff. 2017. UCI Machine Learning Repository. Retrieved from http:\/\/archive.ics.uci.edu\/ml."},{"key":"e_1_2_1_26_1","volume-title":"Adaptive subgradient methods for online learning and stochastic optimization. Journal of Machine Learning Research 12(Jul","author":"Duchi John","year":"2011","unstructured":"John Duchi , Elad Hazan , and Yoram Singer . 2011. Adaptive subgradient methods for online learning and stochastic optimization. Journal of Machine Learning Research 12(Jul 2011 ), 2121--2159. John Duchi, Elad Hazan, and Yoram Singer. 2011. Adaptive subgradient methods for online learning and stochastic optimization. Journal of Machine Learning Research 12(Jul 2011), 2121--2159."},{"key":"e_1_2_1_27_1","volume-title":"Proceedings of the 36th International Conference on Machine Learning (ICML'19)","volume":"97","author":"Eichner Hubert","year":"2019","unstructured":"Hubert Eichner , Tomer Koren , Brendan McMahan , Nathan Srebro , and Kunal Talwar . 2019 . Semi-cyclic stochastic gradient descent . In Proceedings of the 36th International Conference on Machine Learning (ICML'19) , Vol. 97 . 1764--1773. Hubert Eichner, Tomer Koren, Brendan McMahan, Nathan Srebro, and Kunal Talwar. 2019. Semi-cyclic stochastic gradient descent. In Proceedings of the 36th International Conference on Machine Learning (ICML'19), Vol. 97. 1764--1773."},{"key":"e_1_2_1_28_1","volume-title":"Retrieved","year":"2017","unstructured":"Facebook. 2017 . PyTorch . Retrieved January 20, 2020, from https:\/\/pytorch.org\/. Facebook. 2017. PyTorch. Retrieved January 20, 2020, from https:\/\/pytorch.org\/."},{"key":"e_1_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1145\/3241539.3241559"},{"key":"e_1_2_1_30_1","first-page":"1663","article-title":"Consensus-based distributed support vector machines","author":"Forero Pedro A.","year":"2010","unstructured":"Pedro A. Forero , Alfonso Cano , and Georgios B. Giannakis . 2010 . Consensus-based distributed support vector machines . Journal of Machine Learning Research 11 , May (2010), 1663 -- 1707 . Pedro A. Forero, Alfonso Cano, and Georgios B. Giannakis. 2010. Consensus-based distributed support vector machines. Journal of Machine Learning Research 11, May (2010), 1663--1707.","journal-title":"Journal of Machine Learning Research 11"},{"key":"e_1_2_1_31_1","volume-title":"Mitigating sybils in federated learning poisoning. arXiv preprint arXiv:1808.04866","author":"Fung Clement","year":"2018","unstructured":"Clement Fung , Chris J. M. Yoon , and Ivan Beschastnikh . 2018. Mitigating sybils in federated learning poisoning. arXiv preprint arXiv:1808.04866 ( 2018 ). Clement Fung, Chris J. M. Yoon, and Ivan Beschastnikh. 2018. Mitigating sybils in federated learning poisoning. arXiv preprint arXiv:1808.04866 (2018)."},{"key":"e_1_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1145\/3081333.3081358"},{"key":"e_1_2_1_33_1","volume-title":"2011 IEEE 27th International Conference on Data Engineering (ICDE'11)","author":"Ghoting A.","unstructured":"A. Ghoting , R. Krishnamurthy , E. Pednault , B. Reinwald , V. Sindhwani , S. Tatikonda , Y. Tian , and S. Vaithyanathan . 2011. SystemML: Declarative machine learning on MapReduce . In 2011 IEEE 27th International Conference on Data Engineering (ICDE'11) . 231--242. A. Ghoting, R. Krishnamurthy, E. Pednault, B. Reinwald, V. Sindhwani, S. Tatikonda, Y. Tian, and S. Vaithyanathan. 2011. SystemML: Declarative machine learning on MapReduce. In 2011 IEEE 27th International Conference on Data Engineering (ICDE'11). 231--242."},{"key":"e_1_2_1_34_1","volume-title":"Retrieved","author":"Gibiansky Andrew","year":"2017","unstructured":"Andrew Gibiansky . 2017 . Bringing HPC Techniques to Deep Learning . Retrieved June 10, 2020, from https:\/\/andrew.gibiansky.com\/blog\/machine-learning\/baidu-allreduce\/. Andrew Gibiansky. 2017. Bringing HPC Techniques to Deep Learning. Retrieved June 10, 2020, from https:\/\/andrew.gibiansky.com\/blog\/machine-learning\/baidu-allreduce\/."},{"key":"e_1_2_1_35_1","volume-title":"Retrieved","year":"2017","unstructured":"Google. 2017 . TensorFlow Lite . Retrieved January 20, 2020, from https:\/\/www.tensorflow.org\/lite\/. Google. 2017. TensorFlow Lite. Retrieved January 20, 2020, from https:\/\/www.tensorflow.org\/lite\/."},{"key":"e_1_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.jnca.2018.05.003"},{"key":"e_1_2_1_37_1","volume-title":"Federated learning for mobile keyboard prediction. arXiv preprint arXiv:1811.03604","author":"Hard Andrew","year":"2018","unstructured":"Andrew Hard , Kanishka Rao , Rajiv Mathews , Swaroop Ramaswamy , Fran\u00e7oise Beaufays , Sean Augenstein , Hubert Eichner , Chlo\u00e9 Kiddon , and Daniel Ramage . 2018. Federated learning for mobile keyboard prediction. arXiv preprint arXiv:1811.03604 ( 2018 ). Andrew Hard, Kanishka Rao, Rajiv Mathews, Swaroop Ramaswamy, Fran\u00e7oise Beaufays, Sean Augenstein, Hubert Eichner, Chlo\u00e9 Kiddon, and Daniel Ramage. 2018. Federated learning for mobile keyboard prediction. arXiv preprint arXiv:1811.03604 (2018)."},{"key":"e_1_2_1_38_1","volume-title":"COLA: Communication-efficient decentralized linear learning. arXiv preprint arXiv:1808.04883","author":"He Lie","year":"2018","unstructured":"Lie He , An Bian , and Martin Jaggi . 2018 . COLA: Communication-efficient decentralized linear learning. arXiv preprint arXiv:1808.04883 (2018). Lie He, An Bian, and Martin Jaggi. 2018. COLA: Communication-efficient decentralized linear learning. arXiv preprint arXiv:1808.04883 (2018)."},{"key":"e_1_2_1_39_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.tics.2007.09.004"},{"key":"e_1_2_1_40_1","volume-title":"Phillip B. Gibbons, Garth A. Gibson, Greg Ganger, and Eric P. Xing.","author":"Ho Qirong","year":"2013","unstructured":"Qirong Ho , James Cipar , Henggang Cui , Seunghak Lee , Jin Kyu Kim , Phillip B. Gibbons, Garth A. Gibson, Greg Ganger, and Eric P. Xing. 2013 . More effective distributed ML via a stale synchronous parallel parameter server. In Advances in Neural Information Processing Systems (NeurIPS '13). 1223--1231. Qirong Ho, James Cipar, Henggang Cui, Seunghak Lee, Jin Kyu Kim, Phillip B. Gibbons, Garth A. Gibson, Greg Ganger, and Eric P. Xing. 2013. More effective distributed ML via a stale synchronous parallel parameter server. In Advances in Neural Information Processing Systems (NeurIPS'13). 1223--1231."},{"key":"e_1_2_1_41_1","doi-asserted-by":"publisher","DOI":"10.1145\/3081333.3081360"},{"key":"e_1_2_1_42_1","volume-title":"Secure federated learning in 5G mobile networks. arXiv preprint arXiv:2004.06700","author":"Isaksson Martin","year":"2020","unstructured":"Martin Isaksson and Karl Norrman . 2020. Secure federated learning in 5G mobile networks. arXiv preprint arXiv:2004.06700 ( 2020 ). Martin Isaksson and Karl Norrman. 2020. Secure federated learning in 5G mobile networks. arXiv preprint arXiv:2004.06700 (2020)."},{"key":"e_1_2_1_43_1","volume-title":"Jordan","author":"Jaggi Martin","year":"2014","unstructured":"Martin Jaggi , Virginia Smith , Martin Tak\u00e1c , Jonathan Terhorst , Sanjay Krishnan , Thomas Hofmann , and Michael I . Jordan . 2014 . Communication-efficient distributed dual coordinate ascent. In Advances in Neural Information Processing Systems (NeurIPS '14). 3068--3076. Martin Jaggi, Virginia Smith, Martin Tak\u00e1c, Jonathan Terhorst, Sanjay Krishnan, Thomas Hofmann, and Michael I. Jordan. 2014. Communication-efficient distributed dual coordinate ascent. In Advances in Neural Information Processing Systems (NeurIPS'14). 3068--3076."},{"key":"e_1_2_1_44_1","volume-title":"Proceedings of Machine Learning and Systems 2020 (MLSys'20)","volume":"2","author":"Jiang Xiaotang","year":"2020","unstructured":"Xiaotang Jiang , Huan Wang , Yiliu Chen , Ziqi Wu , Lichuan Wang , Bin Zou , Yafeng Yang , Zongyang Cui , Yu Cai , Tianhang Yu , Chengfei Lv , and Zhihua Wu . 2020 . MNN: A universal and efficient inference engine . In Proceedings of Machine Learning and Systems 2020 (MLSys'20) , Vol. 2 . 1--13. Xiaotang Jiang, Huan Wang, Yiliu Chen, Ziqi Wu, Lichuan Wang, Bin Zou, Yafeng Yang, Zongyang Cui, Yu Cai, Tianhang Yu, Chengfei Lv, and Zhihua Wu. 2020. MNN: A universal and efficient inference engine. In Proceedings of Machine Learning and Systems 2020 (MLSys'20), Vol. 2. 1--13."},{"key":"e_1_2_1_45_1","doi-asserted-by":"publisher","DOI":"10.1145\/1150402.1150429"},{"key":"e_1_2_1_46_1","unstructured":"Rie Johnson and Tong Zhang. 2013. Accelerating stochastic gradient descent using predictive variance reduction. In Advances in Neural Information Processing Systems (NeurIPS'13). 315--323.  Rie Johnson and Tong Zhang. 2013. Accelerating stochastic gradient descent using predictive variance reduction. In Advances in Neural Information Processing Systems (NeurIPS'13). 315--323."},{"key":"e_1_2_1_47_1","volume-title":"et\u00a0al","author":"Kairouz Peter","year":"2019","unstructured":"Peter Kairouz , H. Brendan McMahan , Brendan Avent , et\u00a0al . 2019 . Advances and open problems in federated learning. arXiv preprint arXiv:1912.04977 (2019). Peter Kairouz, H. Brendan McMahan, Brendan Avent, et\u00a0al. 2019. Advances and open problems in federated learning. arXiv preprint arXiv:1912.04977 (2019)."},{"key":"e_1_2_1_48_1","doi-asserted-by":"publisher","DOI":"10.1109\/JIOT.2019.2940820"},{"key":"e_1_2_1_49_1","doi-asserted-by":"publisher","DOI":"10.1109\/MWC.001.1900119"},{"key":"e_1_2_1_50_1","volume-title":"Kingma and Jimmy Ba","author":"Diederik","year":"2014","unstructured":"Diederik P. Kingma and Jimmy Ba . 2014 . Adam : A method for stochastic optimization. arXiv preprint arXiv:1412.6980 (2014). Diederik P. Kingma and Jimmy Ba. 2014. Adam: A method for stochastic optimization. arXiv preprint arXiv:1412.6980 (2014)."},{"key":"e_1_2_1_51_1","volume-title":"Decentralized deep learning with arbitrary communication compression. arXiv preprint arXiv:1907.09356","author":"Koloskova Anastasia","year":"2019","unstructured":"Anastasia Koloskova , Tao Lin , Sebastian U. Stich , and Martin Jaggi . 2019. Decentralized deep learning with arbitrary communication compression. arXiv preprint arXiv:1907.09356 ( 2019 ). Anastasia Koloskova, Tao Lin, Sebastian U. Stich, and Martin Jaggi. 2019. Decentralized deep learning with arbitrary communication compression. arXiv preprint arXiv:1907.09356 (2019)."},{"key":"e_1_2_1_52_1","volume-title":"Decentralized stochastic optimization and gossip algorithms with compressed communication. arXiv preprint arXiv:1902.00340","author":"Koloskova Anastasia","year":"2019","unstructured":"Anastasia Koloskova , Sebastian U. Stich , and Martin Jaggi . 2019. Decentralized stochastic optimization and gossip algorithms with compressed communication. arXiv preprint arXiv:1902.00340 ( 2019 ). Anastasia Koloskova, Sebastian U. Stich, and Martin Jaggi. 2019. Decentralized stochastic optimization and gossip algorithms with compressed communication. arXiv preprint arXiv:1902.00340 (2019)."},{"key":"e_1_2_1_53_1","volume-title":"Stochastic, distributed and federated optimization for machine learning. arXiv preprint arXiv:1707.01155","author":"Konecn\u1ef3 Jakub","year":"2017","unstructured":"Jakub Konecn\u1ef3 . 2017. Stochastic, distributed and federated optimization for machine learning. arXiv preprint arXiv:1707.01155 ( 2017 ). Jakub Konecn\u1ef3. 2017. Stochastic, distributed and federated optimization for machine learning. arXiv preprint arXiv:1707.01155 (2017)."},{"key":"e_1_2_1_54_1","volume-title":"Federated optimization: Distributed optimization beyond the datacenter. arXiv preprint arXiv:1511.03575","author":"Kone\u010dn\u1ef3 Jakub","year":"2015","unstructured":"Jakub Kone\u010dn\u1ef3 , Brendan McMahan , and Daniel Ramage . 2015. Federated optimization: Distributed optimization beyond the datacenter. arXiv preprint arXiv:1511.03575 ( 2015 ). Jakub Kone\u010dn\u1ef3, Brendan McMahan, and Daniel Ramage. 2015. Federated optimization: Distributed optimization beyond the datacenter. arXiv preprint arXiv:1511.03575 (2015)."},{"key":"e_1_2_1_55_1","volume-title":"Federated optimization: Distributed machine learning for on-device intelligence. arXiv preprint arXiv:1610.02527","author":"Kone\u010dn\u1ef3 Jakub","year":"2016","unstructured":"Jakub Kone\u010dn\u1ef3 , H. Brendan McMahan , Daniel Ramage , and Peter Richt\u00e1rik . 2016. Federated optimization: Distributed machine learning for on-device intelligence. arXiv preprint arXiv:1610.02527 ( 2016 ). Jakub Kone\u010dn\u1ef3, H. Brendan McMahan, Daniel Ramage, and Peter Richt\u00e1rik. 2016. Federated optimization: Distributed machine learning for on-device intelligence. arXiv preprint arXiv:1610.02527 (2016)."},{"key":"e_1_2_1_56_1","volume-title":"Ananda Theertha Suresh, and Dave Bacon","author":"Kone\u010dn\u1ef3 Jakub","year":"2016","unstructured":"Jakub Kone\u010dn\u1ef3 , H. Brendan McMahan , Felix X. Yu , Peter Richt\u00e1rik , Ananda Theertha Suresh, and Dave Bacon . 2016 . Federated learning: Strategies for improving communication efficiency. arXiv preprint arXiv:1610.05492 (2016). Jakub Kone\u010dn\u1ef3, H. Brendan McMahan, Felix X. Yu, Peter Richt\u00e1rik, Ananda Theertha Suresh, and Dave Bacon. 2016. Federated learning: Strategies for improving communication efficiency. arXiv preprint arXiv:1610.05492 (2016)."},{"key":"e_1_2_1_57_1","doi-asserted-by":"publisher","DOI":"10.1145\/2750858.2804262"},{"key":"e_1_2_1_58_1","volume-title":"Deep learning. Nature 521, 7553","author":"LeCun Yann","year":"2015","unstructured":"Yann LeCun , Yoshua Bengio , and Geoffrey Hinton . 2015. Deep learning. Nature 521, 7553 ( 2015 ), 436. Yann LeCun, Yoshua Bengio, and Geoffrey Hinton. 2015. Deep learning. Nature 521, 7553 (2015), 436."},{"key":"e_1_2_1_59_1","doi-asserted-by":"publisher","DOI":"10.1109\/5.726791"},{"key":"e_1_2_1_60_1","article-title":"Rcv1: A new benchmark collection for text categorization research","author":"Lewis David D.","year":"2004","unstructured":"David D. Lewis , Yiming Yang , Tony G. Rose , and Fan Li . 2004 . Rcv1: A new benchmark collection for text categorization research . Journal of Machine Learning Research 5 (Apr 2004), 361--397. David D. Lewis, Yiming Yang, Tony G. Rose, and Fan Li. 2004. Rcv1: A new benchmark collection for text categorization research. Journal of Machine Learning Research 5 (Apr 2004), 361--397.","journal-title":"Journal of Machine Learning Research 5"},{"key":"e_1_2_1_61_1","doi-asserted-by":"publisher","DOI":"10.1145\/2640087.2644155"},{"key":"e_1_2_1_62_1","volume-title":"Ameet Talwalkar, and Virginia Smith.","author":"Li Tian","year":"2019","unstructured":"Tian Li , Anit Kumar Sahu , Ameet Talwalkar, and Virginia Smith. 2019 . Federated learning: Challenges , methods, and future directions. arXiv preprint arXiv:1908.07873 (2019). Tian Li, Anit Kumar Sahu, Ameet Talwalkar, and Virginia Smith. 2019. Federated learning: Challenges, methods, and future directions. arXiv preprint arXiv:1908.07873 (2019)."},{"key":"e_1_2_1_63_1","unstructured":"Xiangru Lian Yijun Huang Yuncheng Li and Ji Liu. 2015. Asynchronous parallel stochastic gradient for nonconvex optimization. In Advances in Neural Information Processing Systems (NeurIPS'15). 2737--2745.  Xiangru Lian Yijun Huang Yuncheng Li and Ji Liu. 2015. Asynchronous parallel stochastic gradient for nonconvex optimization. In Advances in Neural Information Processing Systems (NeurIPS'15). 2737--2745."},{"key":"e_1_2_1_64_1","volume-title":"Secure federated transfer learning. arXiv preprint arXiv:1812.03337","author":"Liu Yang","year":"2018","unstructured":"Yang Liu , Tianjian Chen , and Qiang Yang . 2018. Secure federated transfer learning. arXiv preprint arXiv:1812.03337 ( 2018 ). Yang Liu, Tianjian Chen, and Qiang Yang. 2018. Secure federated transfer learning. arXiv preprint arXiv:1812.03337 (2018)."},{"key":"e_1_2_1_65_1","doi-asserted-by":"publisher","DOI":"10.1109\/TKDE.2020.2967670"},{"key":"e_1_2_1_66_1","volume-title":"Adding vs. averaging in distributed primal-dual optimization. arXiv preprint arXiv:1502.03508","author":"Ma Chenxin","year":"2015","unstructured":"Chenxin Ma , Virginia Smith , Martin Jaggi , Michael I. Jordan , Peter Richt\u00e1rik , and Martin Tak\u00e1\u010d . 2015. Adding vs. averaging in distributed primal-dual optimization. arXiv preprint arXiv:1502.03508 ( 2015 ). Chenxin Ma, Virginia Smith, Martin Jaggi, Michael I. Jordan, Peter Richt\u00e1rik, and Martin Tak\u00e1\u010d. 2015. Adding vs. averaging in distributed primal-dual optimization. arXiv preprint arXiv:1502.03508 (2015)."},{"key":"e_1_2_1_67_1","volume-title":"Distributed inexact damped newton method: Data partitioning and load-balancing. arXiv preprint arXiv:1603.05191","author":"Ma Chenxin","year":"2016","unstructured":"Chenxin Ma and Martin Tak\u00e1\u010d . 2016. Distributed inexact damped newton method: Data partitioning and load-balancing. arXiv preprint arXiv:1603.05191 ( 2016 ). Chenxin Ma and Martin Tak\u00e1\u010d. 2016. Distributed inexact damped newton method: Data partitioning and load-balancing. arXiv preprint arXiv:1603.05191 (2016)."},{"key":"e_1_2_1_68_1","doi-asserted-by":"publisher","DOI":"10.1145\/3081333.3081359"},{"key":"e_1_2_1_69_1","volume-title":"2016 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP\u201916)","author":"McGraw I.","unstructured":"I. McGraw , R. Prabhavalkar , R. Alvarez , M. G. Arenas , K. Rao , D. Rybach , O. Alsharif , H. Sak , A. Gruenstein , F. Beaufays , and C. Parada . 2016. Personalized speech recognition on mobile devices . In 2016 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP\u201916) . 5955--5959. I. McGraw, R. Prabhavalkar, R. Alvarez, M. G. Arenas, K. Rao, D. Rybach, O. Alsharif, H. Sak, A. Gruenstein, F. Beaufays, and C. Parada. 2016. Personalized speech recognition on mobile devices. In 2016 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP\u201916). 5955--5959."},{"key":"e_1_2_1_70_1","volume-title":"Communication-efficient learning of deep networks from decentralized data. arXiv preprint arXiv:1602.05629","author":"McMahan H. Brendan","year":"2016","unstructured":"H. Brendan McMahan , Eider Moore , Daniel Ramage , and Blaise Ag\u00fcera y Arcas . 2016. Communication-efficient learning of deep networks from decentralized data. arXiv preprint arXiv:1602.05629 ( 2016 ). H. Brendan McMahan, Eider Moore, Daniel Ramage, and Blaise Ag\u00fcera y Arcas. 2016. Communication-efficient learning of deep networks from decentralized data. arXiv preprint arXiv:1602.05629 (2016)."},{"key":"e_1_2_1_71_1","volume-title":"Learning differentially private language models without losing accuracy. arXiv preprint arXiv:1710.06963","author":"McMahan H. Brendan","year":"2017","unstructured":"H. Brendan McMahan , Daniel Ramage , Kunal Talwar , and Li Zhang . 2017. Learning differentially private language models without losing accuracy. arXiv preprint arXiv:1710.06963 ( 2017 ). H. Brendan McMahan, Daniel Ramage, Kunal Talwar, and Li Zhang. 2017. Learning differentially private language models without losing accuracy. arXiv preprint arXiv:1710.06963 (2017)."},{"key":"e_1_2_1_72_1","first-page":"1","article-title":"MLlib: Machine learning in apache spark","volume":"17","author":"Meng Xiangrui","year":"2016","unstructured":"Xiangrui Meng , Joseph Bradley , Burak Yavuz , Evan Sparks , Shivaram Venkataraman , Davies Liu , Jeremy Freeman , D. B. Tsai , Manish Amde , Sean Owen , Doris Xin , Reynold Xin , Michael J. Franklin , Reza Zadeh , Matei Zaharia , and Ameet Talwalkar . 2016 . MLlib: Machine learning in apache spark . Journal of Machine Learning Research 17 , 34 (2016), 1 -- 7 . Xiangrui Meng, Joseph Bradley, Burak Yavuz, Evan Sparks, Shivaram Venkataraman, Davies Liu, Jeremy Freeman, D. B. Tsai, Manish Amde, Sean Owen, Doris Xin, Reynold Xin, Michael J. Franklin, Reza Zadeh, Matei Zaharia, and Ameet Talwalkar. 2016. MLlib: Machine learning in apache spark. Journal of Machine Learning Research 17, 34 (2016), 1--7.","journal-title":"Journal of Machine Learning Research"},{"key":"e_1_2_1_73_1","volume-title":"Proceedings of the 2016 ACM International Joint Conference on Pervasive and Ubiquitous Computing (UbiComp'16)","author":"Mittal Gaurav","unstructured":"Gaurav Mittal , Kaushal B. Yagnik , Mohit Garg , and Narayanan C. Krishnan . 2016. SpotGarbage: Smartphone app to detect garbage using deep learning . In Proceedings of the 2016 ACM International Joint Conference on Pervasive and Ubiquitous Computing (UbiComp'16) . 940--945. Gaurav Mittal, Kaushal B. Yagnik, Mohit Garg, and Narayanan C. Krishnan. 2016. SpotGarbage: Smartphone app to detect garbage using deep learning. In Proceedings of the 2016 ACM International Joint Conference on Pervasive and Ubiquitous Computing (UbiComp'16). 940--945."},{"key":"e_1_2_1_74_1","volume-title":"Proceedings of the 36th International Conference on Machine Learning (ICML'19)","volume":"97","author":"Mohri Mehryar","year":"2019","unstructured":"Mehryar Mohri , Gary Sivek , and Ananda Theertha Suresh . 2019 . Agnostic federated learning . In Proceedings of the 36th International Conference on Machine Learning (ICML'19) , Vol. 97 . 4615--4625. Mehryar Mohri, Gary Sivek, and Ananda Theertha Suresh. 2019. Agnostic federated learning. In Proceedings of the 36th International Conference on Machine Learning (ICML'19), Vol. 97. 4615--4625."},{"key":"e_1_2_1_75_1","volume-title":"Reed","author":"Niknam Solmaz","year":"2019","unstructured":"Solmaz Niknam , Harpreet S. Dhillon , and Jeffery H . Reed . 2019 . Federated learning for wireless communications: Motivation , opportunities and challenges. arXiv preprint arXiv:1908.06847 (2019). Solmaz Niknam, Harpreet S. Dhillon, and Jeffery H. Reed. 2019. Federated learning for wireless communications: Motivation, opportunities and challenges. arXiv preprint arXiv:1908.06847 (2019)."},{"key":"e_1_2_1_76_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.jpdc.2008.09.002"},{"key":"e_1_2_1_77_1","doi-asserted-by":"publisher","DOI":"10.1007\/s13748-012-0035-5"},{"key":"e_1_2_1_78_1","volume-title":"Split learning for collaborative deep learning in healthcare. arXiv preprint arXiv:1912.12115","author":"Poirot Maarten G.","year":"2019","unstructured":"Maarten G. Poirot , Praneeth Vepakomma , Ken Chang , Jayashree Kalpathy-Cramer , Rajiv Gupta , and Ramesh Raskar . 2019. Split learning for collaborative deep learning in healthcare. arXiv preprint arXiv:1912.12115 ( 2019 ). Maarten G. Poirot, Praneeth Vepakomma, Ken Chang, Jayashree Kalpathy-Cramer, Rajiv Gupta, and Ramesh Raskar. 2019. Split learning for collaborative deep learning in healthcare. arXiv preprint arXiv:1912.12115 (2019)."},{"key":"e_1_2_1_80_1","doi-asserted-by":"publisher","DOI":"10.1145\/2043556.2043566"},{"key":"e_1_2_1_81_1","doi-asserted-by":"publisher","DOI":"10.1145\/2330667.2330691"},{"key":"e_1_2_1_82_1","volume-title":"Proceedings of the Thirteenth National Conference on Artificial Intelligence and Eighth Innovative Applications of Artificial Intelligence Conference (AAAI\/IAAI'96)","volume":"1","author":"Foster","unstructured":"Foster J. Provost and Daniel N. Hennessy. 1996. Scaling up: Distributed machine learning with cooperation . In Proceedings of the Thirteenth National Conference on Artificial Intelligence and Eighth Innovative Applications of Artificial Intelligence Conference (AAAI\/IAAI'96) , Vol. 1 . 74--79. Foster J. Provost and Daniel N. Hennessy. 1996. Scaling up: Distributed machine learning with cooperation. In Proceedings of the Thirteenth National Conference on Artificial Intelligence and Eighth Innovative Applications of Artificial Intelligence Conference (AAAI\/IAAI'96), Vol. 1. 74--79."},{"key":"e_1_2_1_83_1","doi-asserted-by":"publisher","DOI":"10.1016\/S0893-6080(98)00116-6"},{"key":"e_1_2_1_84_1","volume-title":"IEEE IEEE Conference on Computer Communications (INFOCOM\u201918)","author":"Ran X.","unstructured":"X. Ran , H. Chen , X. Zhu , Z. Liu , and J. Chen . 2018. DeepDecision: A mobile deep learning framework for edge video analytics . In IEEE IEEE Conference on Computer Communications (INFOCOM\u201918) . 1421--1429. X. Ran, H. Chen, X. Zhu, Z. Liu, and J. Chen. 2018. DeepDecision: A mobile deep learning framework for edge video analytics. In IEEE IEEE Conference on Computer Communications (INFOCOM\u201918). 1421--1429."},{"key":"e_1_2_1_85_1","volume-title":"Proceedings of the 36th International Conference on Machine Learning (ICML'19)","volume":"97","author":"Ravi Sujith","year":"2019","unstructured":"Sujith Ravi . 2019 . Efficient on-device models using neural projections . In Proceedings of the 36th International Conference on Machine Learning (ICML'19) , Vol. 97 . 5370--5379. Sujith Ravi. 2019. Efficient on-device models using neural projections. In Proceedings of the 36th International Conference on Machine Learning (ICML'19), Vol. 97. 5370--5379."},{"key":"e_1_2_1_86_1","volume-title":"Hogwild: A lock-free approach to parallelizing stochastic gradient descent. In Advances in Neural Information Processing Systems (NeurIPS'11). 693--701.","author":"Recht Benjamin","year":"2011","unstructured":"Benjamin Recht , Christopher Re , Stephen Wright , and Feng Niu . 2011 . Hogwild: A lock-free approach to parallelizing stochastic gradient descent. In Advances in Neural Information Processing Systems (NeurIPS'11). 693--701. Benjamin Recht, Christopher Re, Stephen Wright, and Feng Niu. 2011. Hogwild: A lock-free approach to parallelizing stochastic gradient descent. In Advances in Neural Information Processing Systems (NeurIPS'11). 693--701."},{"key":"e_1_2_1_87_1","volume-title":"Retrieved","author":"Rennie Jason","year":"2007","unstructured":"Jason Rennie . 2007 . 20 Newsgroups . Retrieved June 22, 2020, from http:\/\/qwone.com\/ jason\/20Newsgroups\/. Jason Rennie. 2007. 20 Newsgroups. Retrieved June 22, 2020, from http:\/\/qwone.com\/ jason\/20Newsgroups\/."},{"key":"e_1_2_1_88_1","doi-asserted-by":"publisher","DOI":"10.5555\/2946645.3007028"},{"key":"e_1_2_1_89_1","doi-asserted-by":"publisher","DOI":"10.1214\/aoms\/1177729586"},{"key":"e_1_2_1_90_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-015-0816-y"},{"key":"e_1_2_1_91_1","doi-asserted-by":"publisher","DOI":"10.1109\/GLOCOM.2018.8647927"},{"key":"e_1_2_1_92_1","volume-title":"Proceedings of the 6th International Conference on Learning Representations (ICLR'18)","author":"Sashank J. Reddi","year":"2018","unstructured":"J. Reddi Sashank , Kale Satyen , and Kumar Sanjiv . 2018 . On the convergence of Adam and beyond . In Proceedings of the 6th International Conference on Learning Representations (ICLR'18) . J. Reddi Sashank, Kale Satyen, and Kumar Sanjiv. 2018. On the convergence of Adam and beyond. In Proceedings of the 6th International Conference on Learning Representations (ICLR'18)."},{"key":"e_1_2_1_93_1","volume-title":"Robust and communication-efficient federated learning from non-IID data. arXiv preprint arXiv:1903.02891","author":"Sattler Felix","year":"2019","unstructured":"Felix Sattler , Simon Wiedemann , Klaus-Robert M\u00fcller , and Wojciech Samek . 2019. Robust and communication-efficient federated learning from non-IID data. arXiv preprint arXiv:1903.02891 ( 2019 ). Felix Sattler, Simon Wiedemann, Klaus-Robert M\u00fcller, and Wojciech Samek. 2019. Robust and communication-efficient federated learning from non-IID data. arXiv preprint arXiv:1903.02891 (2019)."},{"key":"e_1_2_1_94_1","doi-asserted-by":"publisher","DOI":"10.1109\/JIOT.2020.2964162"},{"key":"e_1_2_1_95_1","volume-title":"Horovod: Fast and easy distributed deep learning in TensorFlow. arXiv preprint arXiv:1802.05799","author":"Sergeev Alexander","year":"2018","unstructured":"Alexander Sergeev and Mike Del Balso . 2018 . Horovod: Fast and easy distributed deep learning in TensorFlow. arXiv preprint arXiv:1802.05799 (2018). Alexander Sergeev and Mike Del Balso. 2018. Horovod: Fast and easy distributed deep learning in TensorFlow. arXiv preprint arXiv:1802.05799 (2018)."},{"key":"e_1_2_1_96_1","volume-title":"Proceedings of the 31st International Conference on Machine Learning (ICML'14)","volume":"32","author":"Shamir Ohad","year":"2014","unstructured":"Ohad Shamir , Nati Srebro , and Tong Zhang . 2014 . Communication-efficient distributed optimization using an approximate newton-type method . In Proceedings of the 31st International Conference on Machine Learning (ICML'14) , Vol. 32 . 1000--1008. Ohad Shamir, Nati Srebro, and Tong Zhang. 2014. Communication-efficient distributed optimization using an approximate newton-type method. In Proceedings of the 31st International Conference on Machine Learning (ICML'14), Vol. 32. 1000--1008."},{"key":"e_1_2_1_97_1","volume-title":"ExpertMatcher: Automating ML model selection for clients using hidden representations. arXiv preprint arXiv:1910.03731","author":"Sharma Vivek","year":"2019","unstructured":"Vivek Sharma , Praneeth Vepakomma , Tristan Swedish , Ken Chang , Jayashree Kalpathy-Cramer , and Ramesh Raskar . 2019. ExpertMatcher: Automating ML model selection for clients using hidden representations. arXiv preprint arXiv:1910.03731 ( 2019 ). Vivek Sharma, Praneeth Vepakomma, Tristan Swedish, Ken Chang, Jayashree Kalpathy-Cramer, and Ramesh Raskar. 2019. ExpertMatcher: Automating ML model selection for clients using hidden representations. arXiv preprint arXiv:1910.03731 (2019)."},{"key":"e_1_2_1_98_1","volume-title":"Biscotti: A ledger for private and secure peer-to-peer machine learning. arXiv preprint arXiv:1811.09904","author":"Shayan Muhammad","year":"2018","unstructured":"Muhammad Shayan , Clement Fung , Chris J. M. Yoon , and Ivan Beschastnikh . 2018 . Biscotti: A ledger for private and secure peer-to-peer machine learning. arXiv preprint arXiv:1811.09904 (2018). Muhammad Shayan, Clement Fung, Chris J. M. Yoon, and Ivan Beschastnikh. 2018. Biscotti: A ledger for private and secure peer-to-peer machine learning. arXiv preprint arXiv:1811.09904 (2018)."},{"key":"e_1_2_1_99_1","volume-title":"International MICCAI Brainlesion Workshop. 92--104","author":"Sheller Micah J.","year":"2018","unstructured":"Micah J. Sheller , G. Anthony Reina , Brandon Edwards , Jason Martin , and Spyridon Bakas . 2018 . Multi-institutional deep learning modeling without sharing patient data: A feasibility study on brain tumor segmentation . In International MICCAI Brainlesion Workshop. 92--104 . Micah J. Sheller, G. Anthony Reina, Brandon Edwards, Jason Martin, and Spyridon Bakas. 2018. Multi-institutional deep learning modeling without sharing patient data: A feasibility study on brain tumor segmentation. In International MICCAI Brainlesion Workshop. 92--104."},{"key":"e_1_2_1_100_1","doi-asserted-by":"publisher","DOI":"10.1145\/2810103.2813687"},{"key":"e_1_2_1_101_1","doi-asserted-by":"publisher","DOI":"10.1109\/SP.2017.41"},{"key":"e_1_2_1_102_1","doi-asserted-by":"publisher","DOI":"10.1145\/3090082"},{"key":"e_1_2_1_103_1","volume-title":"Detailed comparison of communication efficiency of split learning and federated learning. arXiv preprint arXiv:1909.09145","author":"Singh Abhishek","year":"2019","unstructured":"Abhishek Singh , Praneeth Vepakomma , Otkrist Gupta , and Ramesh Raskar . 2019. Detailed comparison of communication efficiency of split learning and federated learning. arXiv preprint arXiv:1909.09145 ( 2019 ). Abhishek Singh, Praneeth Vepakomma, Otkrist Gupta, and Ramesh Raskar. 2019. Detailed comparison of communication efficiency of split learning and federated learning. arXiv preprint arXiv:1909.09145 (2019)."},{"key":"e_1_2_1_104_1","volume-title":"Retrieved","author":"Team Siri","year":"2017","unstructured":"Siri Team . 2017 . Deep learning for Siri\u2019s voice: On-device deep mixture density networks for hybrid unit selection synthesis . Retrieved Nov. 13, 2020, from https:\/\/machinelearning.apple.com\/research\/siri-voices. Siri Team. 2017. Deep learning for Siri\u2019s voice: On-device deep mixture density networks for hybrid unit selection synthesis. Retrieved Nov. 13, 2020, from https:\/\/machinelearning.apple.com\/research\/siri-voices."},{"key":"e_1_2_1_105_1","volume-title":"Advances in Neural Information Processing Systems 30 (NeurIPS'17). Curran Associates","author":"Smith Virginia","unstructured":"Virginia Smith , Chao-Kai Chiang , Maziar Sanjabi , and Ameet S Talwalkar . 2017. Federated multi-task learning . In Advances in Neural Information Processing Systems 30 (NeurIPS'17). Curran Associates , Inc ., 4424--4434. Virginia Smith, Chao-Kai Chiang, Maziar Sanjabi, and Ameet S Talwalkar. 2017. Federated multi-task learning. In Advances in Neural Information Processing Systems 30 (NeurIPS'17). Curran Associates, Inc., 4424--4434."},{"key":"e_1_2_1_106_1","doi-asserted-by":"publisher","DOI":"10.5555\/3122009.3290415"},{"key":"e_1_2_1_107_1","doi-asserted-by":"publisher","DOI":"10.14778\/1920841.1920931"},{"key":"e_1_2_1_108_1","volume-title":"Retrieved","author":"Stephen Boyd","year":"2010","unstructured":"Boyd Stephen , Parikh Neal , Chu Eric , Peleato Borja , and Eckstein Jonathan . 2010 . MPI example for alternating direction method of multipliers . Retrieved June 26, 2020, from https:\/\/stanford.edu\/ boyd\/papers\/admm\/mpi\/. Boyd Stephen, Parikh Neal, Chu Eric, Peleato Borja, and Eckstein Jonathan. 2010. MPI example for alternating direction method of multipliers. Retrieved June 26, 2020, from https:\/\/stanford.edu\/ boyd\/papers\/admm\/mpi\/."},{"key":"e_1_2_1_109_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-71249-9_12"},{"key":"e_1_2_1_110_1","volume-title":"Retrieved","year":"2017","unstructured":"Tencent. 2017 . ncnn . Retrieved January 20, 2020, from https:\/\/github.com\/Tencent\/ncnn. Tencent. 2017. ncnn. Retrieved January 20, 2020, from https:\/\/github.com\/Tencent\/ncnn."},{"key":"e_1_2_1_111_1","unstructured":"T. Tieleman and Geoffrey Hinton. 2012. Neural networks for machine learning. Retrieved from https:\/\/www.cs.toronto.edu\/tijmen\/csc321\/slides\/lecture_slides_lec6.pdf.  T. Tieleman and Geoffrey Hinton. 2012. Neural networks for machine learning. Retrieved from https:\/\/www.cs.toronto.edu\/tijmen\/csc321\/slides\/lecture_slides_lec6.pdf."},{"key":"e_1_2_1_112_1","volume-title":"Reducing leakage in distributed deep learning for sensitive health data. arXiv preprint arXiv:1812.00564","author":"Vepakomma Praneeth","year":"2019","unstructured":"Praneeth Vepakomma , Otkrist Gupta , Abhimanyu Dubey , and Ramesh Raskar . 2019. Reducing leakage in distributed deep learning for sensitive health data. arXiv preprint arXiv:1812.00564 ( 2019 ). Praneeth Vepakomma, Otkrist Gupta, Abhimanyu Dubey, and Ramesh Raskar. 2019. Reducing leakage in distributed deep learning for sensitive health data. arXiv preprint arXiv:1812.00564 (2019)."},{"key":"e_1_2_1_113_1","volume-title":"Split learning for health: Distributed deep learning without sharing raw patient data. arXiv preprint arXiv:1812.00564","author":"Vepakomma Praneeth","year":"2018","unstructured":"Praneeth Vepakomma , Otkrist Gupta , Tristan Swedish , and Ramesh Raskar . 2018. Split learning for health: Distributed deep learning without sharing raw patient data. arXiv preprint arXiv:1812.00564 ( 2018 ). Praneeth Vepakomma, Otkrist Gupta, Tristan Swedish, and Ramesh Raskar. 2018. Split learning for health: Distributed deep learning without sharing raw patient data. arXiv preprint arXiv:1812.00564 (2018)."},{"key":"e_1_2_1_114_1","volume-title":"No peek: A survey of private distributed deep learning. arXiv preprint arXiv:1812.03288","author":"Vepakomma Praneeth","year":"2018","unstructured":"Praneeth Vepakomma , Tristan Swedish , Ramesh Raskar , Otkrist Gupta , and Abhimanyu Dubey . 2018. No peek: A survey of private distributed deep learning. arXiv preprint arXiv:1812.03288 ( 2018 ). Praneeth Vepakomma, Tristan Swedish, Ramesh Raskar, Otkrist Gupta, and Abhimanyu Dubey. 2018. No peek: A survey of private distributed deep learning. arXiv preprint arXiv:1812.03288 (2018)."},{"key":"e_1_2_1_115_1","volume-title":"Advances in Neural Information Processing Systems 30 (NeurIPS'17). Curran Associates","author":"Wilson Ashia C.","unstructured":"Ashia C. Wilson , Rebecca Roelofs , Mitchell Stern , Nati Srebro , and Benjamin Recht . 2017. The marginal value of adaptive gradient methods in machine learning . In Advances in Neural Information Processing Systems 30 (NeurIPS'17). Curran Associates , Inc ., 4148--4158. Ashia C. Wilson, Rebecca Roelofs, Mitchell Stern, Nati Srebro, and Benjamin Recht. 2017. The marginal value of adaptive gradient methods in machine learning. In Advances in Neural Information Processing Systems 30 (NeurIPS'17). Curran Associates, Inc., 4148--4158."},{"key":"e_1_2_1_116_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10115-007-0114-2"},{"key":"e_1_2_1_117_1","doi-asserted-by":"publisher","DOI":"10.1145\/3308558.3313591"},{"key":"e_1_2_1_118_1","doi-asserted-by":"publisher","DOI":"10.1145\/3298981"},{"key":"e_1_2_1_119_1","volume-title":"Applied federated learning: Improving Google keyboard query suggestions. arXiv preprint arXiv:1812.02903","author":"Yang Timothy","year":"2018","unstructured":"Timothy Yang , Galen Andrew , Hubert Eichner , Haicheng Sun , Wei Li , Nicholas Kong , Daniel Ramage , and Fran\u00e7oise Beaufays . 2018. Applied federated learning: Improving Google keyboard query suggestions. arXiv preprint arXiv:1812.02903 ( 2018 ). Timothy Yang, Galen Andrew, Hubert Eichner, Haicheng Sun, Wei Li, Nicholas Kong, Daniel Ramage, and Fran\u00e7oise Beaufays. 2018. Applied federated learning: Improving Google keyboard query suggestions. arXiv preprint arXiv:1812.02903 (2018)."},{"key":"e_1_2_1_120_1","volume-title":"Proceedings of the 36th International Conference on Machine Learning (ICML'19)","volume":"97","author":"Yu Hao","year":"2019","unstructured":"Hao Yu , Rong Jin , and Sen Yang . 2019 . On the linear speedup analysis of communication efficient momentum SGD for distributed non-convex optimization . In Proceedings of the 36th International Conference on Machine Learning (ICML'19) , Vol. 97 . 7184--7193. Hao Yu, Rong Jin, and Sen Yang. 2019. On the linear speedup analysis of communication efficient momentum SGD for distributed non-convex optimization. In Proceedings of the 36th International Conference on Machine Learning (ICML'19), Vol. 97. 7184--7193."},{"key":"e_1_2_1_121_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v33i01.33015693"},{"key":"e_1_2_1_122_1","doi-asserted-by":"publisher","DOI":"10.1145\/3081333.3081336"},{"key":"e_1_2_1_123_1","unstructured":"Sixin Zhang Anna E. Choromanska and Yann LeCun. 2015. Deep learning with elastic averaging SGD. In Advances in Neural Information Processing Systems (NeurIPS'15). 685--693.  Sixin Zhang Anna E. Choromanska and Yann LeCun. 2015. Deep learning with elastic averaging SGD. In Advances in Neural Information Processing Systems (NeurIPS'15). 685--693."},{"key":"e_1_2_1_124_1","volume-title":"International Conference on Machine Learning (ICML'15)","volume":"37","author":"Zhang Yuchen","year":"2015","unstructured":"Yuchen Zhang and Xiao Lin . 2015 . DiSCO: Distributed optimization for self-concordant empirical loss . In International Conference on Machine Learning (ICML'15) , Vol. 37 . 362--370. Yuchen Zhang and Xiao Lin. 2015. DiSCO: Distributed optimization for self-concordant empirical loss. In International Conference on Machine Learning (ICML'15), Vol. 37. 362--370."},{"key":"e_1_2_1_125_1","volume-title":"Federated learning with non-IID data. arXiv preprint arXiv:1806.00582","author":"Zhao Yue","year":"2018","unstructured":"Yue Zhao , Meng Li , Liangzhen Lai , Naveen Suda , Damon Civin , and Vikas Chandra . 2018. Federated learning with non-IID data. arXiv preprint arXiv:1806.00582 ( 2018 ). Yue Zhao, Meng Li, Liangzhen Lai, Naveen Suda, Damon Civin, and Vikas Chandra. 2018. Federated learning with non-IID data. arXiv preprint arXiv:1806.00582 (2018)."},{"key":"e_1_2_1_126_1","doi-asserted-by":"publisher","DOI":"10.5555\/3305890.3306107"},{"key":"e_1_2_1_127_1","volume-title":"Smola","author":"Zinkevich Martin","year":"2010","unstructured":"Martin Zinkevich , Markus Weimer , Lihong Li , and Alex J . Smola . 2010 . Parallelized stochastic gradient descent. In Advances in Neural Information Processing Systems (NeurIPS '10). 2595--2603. Martin Zinkevich, Markus Weimer, Lihong Li, and Alex J. Smola. 2010. Parallelized stochastic gradient descent. In Advances in Neural Information Processing Systems (NeurIPS'10). 2595--2603."}],"container-title":["ACM Computing Surveys"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3424660","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3424660","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T23:23:31Z","timestamp":1750202611000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3424660"}},"subtitle":["A Comprehensive Survey"],"short-title":[],"issued":{"date-parts":[[2021,1,2]]},"references-count":126,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2022,1,31]]}},"alternative-id":["10.1145\/3424660"],"URL":"https:\/\/doi.org\/10.1145\/3424660","relation":{},"ISSN":["0360-0300","1557-7341"],"issn-type":[{"value":"0360-0300","type":"print"},{"value":"1557-7341","type":"electronic"}],"subject":[],"published":{"date-parts":[[2021,1,2]]},"assertion":[{"value":"2020-01-01","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2020-09-01","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2021-01-02","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}