{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,2]],"date-time":"2026-07-02T05:52:14Z","timestamp":1782971534285,"version":"3.54.5"},"reference-count":191,"publisher":"Informa UK Limited","issue":"8","content-domain":{"domain":["www.tandfonline.com"],"crossmark-restriction":true},"short-container-title":["Journal of Experimental &amp; Theoretical Artificial Intelligence"],"published-print":{"date-parts":[[2025,11,17]]},"DOI":"10.1080\/0952813x.2024.2391778","type":"journal-article","created":{"date-parts":[[2024,9,12]],"date-time":"2024-09-12T07:23:03Z","timestamp":1726125783000},"page":"1331-1389","update-policy":"https:\/\/doi.org\/10.1080\/tandf_crossmark_01","source":"Crossref","is-referenced-by-count":9,"title":["Byzantine fault tolerance in distributed machine learning: a survey"],"prefix":"10.1080","volume":"37","author":[{"given":"Djamila","family":"Bouhata","sequence":"first","affiliation":[{"name":"Computer Science Department, University of Batna 2, Batna, Algeria"},{"name":"Laboratory of Applications of Mathematics to Computer and Electronics","place":["Batna, Algeria"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Hamouma","family":"Moumen","sequence":"additional","affiliation":[{"name":"Computer Science Department, University of Batna 2, Batna, Algeria"},{"name":"Laboratory of Applications of Mathematics to Computer and Electronics","place":["Batna, Algeria"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Jocelyn Ahmed","family":"Mazari","sequence":"additional","affiliation":[{"name":"ISIR, CNRS, Sorbonne Universite","place":["Paris, France"]},{"name":"Extrality","place":["Paris, France"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Ahc\u00e8ne","family":"Bounceur","sequence":"additional","affiliation":[{"name":"University of Sharjah, University City","place":["Sharjah, United Arab Emirates"]}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"301","published-online":{"date-parts":[[2024,9,12]]},"reference":[{"key":"e_1_3_4_2_1","first-page":"7575","volume-title":"Advances in Neural Information Processing Systems","volume":"31","author":"Agarwal N.","year":"2018","unstructured":"Agarwal, N., Suresh, A. T., Yu, F. X., Kumar, S., & McMahan, B. (2018). cpSGD: Communication-efficient and differentially-private distributed SGD. In: S. Bengio, H. Wallach, H. Larochelle, K. Grauman, N. Cesa-Bianchi, & R. Garnett (Eds.). Advances in Neural Information Processing Systems (Vol. 31. pp. 7575\u20137586). Curran Associates, Inc. 32nd Conference on Neural Information Processing Systems (NIPS 2018), Montr\u00e9al, Canada. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2018\/file\/21ce689121e39821d07d04faab328370-Paper.pdf"},{"key":"e_1_3_4_3_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.jcss.2013.02.006"},{"key":"e_1_3_4_4_1","first-page":"4613","volume-title":"Advances in Neural Information Processing Systems","volume":"31","author":"Alistarh D.","year":"2018","unstructured":"Alistarh, D., Allen-Zhu, Z., & Li, J. (2018). Byzantine stochastic gradient descent. In: S. Bengio, H. Wallach, H. Larochelle, K. Grauman, N. Cesa-Bianchi, & R. Garnett (Eds.), Advances in Neural Information Processing Systems (Vol. 31. pp. 4613\u20134623). Curran Associates, Inc. 32nd Conference on Neural Information Processing Systems (NeurIPS 2018), Montreal, Canada. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2018\/file\/a07c2f3b3b907aaf8436a26c6d77f0a2-Paper.pdf"},{"key":"e_1_3_4_5_1","doi-asserted-by":"publisher","DOI":"10.1016\/S0893-6080(96)00089-5"},{"key":"e_1_3_4_6_1","doi-asserted-by":"publisher","DOI":"10.1002\/0471478210"},{"issue":"5","key":"e_1_3_4_7_1","first-page":"1475","article-title":"Crop recommendation system using neural networks","volume":"5","author":"Banavlikar T.","year":"2018","unstructured":"Banavlikar, T., Mahir, A., Budukh, M., & Dhodapkar, S. (2018). Crop recommendation system using neural networks. International Research Journal of Engineering & Technology (IRJET), 5(5), 1475\u20131480. https:\/\/www.irjet.net\/archives\/V5\/i5\/IRJET-V5I5280.pdf","journal-title":"International Research Journal of Engineering & Technology (IRJET)"},{"key":"e_1_3_4_8_1","first-page":"8632","volume-title":"Advances in Neural Information Processing Systems","volume":"32","author":"Baruch G.","year":"2019","unstructured":"Baruch, G., Baruch, M., & Goldberg, Y. (2019). A little is enough: Circumventing defenses for distributed learning. In: H. Wallach, H. Larochelle, A. Beygelzimer, F. d\u2019Alch\u00e9-Buc, E. Fox, & R. Garnett (Eds.). Advances in Neural Information Processing Systems (Vol. 32. pp. 8632\u20138642). Curran Associates, Inc. 33rd Conference on Neural Information Processing Systems (NeurIPS 2019), Vancouver, Canada. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2019\/file\/ec1c59141046cd1866bbbcdfb6ae31d4-Paper.pdf"},{"key":"e_1_3_4_9_1","first-page":"14695","volume-title":"Advances in Neural Information Processing Systems","volume":"32","author":"Basu D.","year":"2019","unstructured":"Basu, D., Data, D., Karakus, C., & Diggavi, S. (2019). Qsparse-local-sgd: Distributed SGD with quantization, sparsification, and local computations. In: H. Wallach, H. Larochelle, A. Beygelzimer, F. d\u2019Alch\u00e9-Buc, E. Fox, & R. Garnett (Eds.). Advances in Neural Information Processing Systems (Vol. 32. pp. 14695\u201314706). Curran Associates, Inc. 33rd Conference on Neural Information Processing Systems (NeurIPS 2019), Vancouver, Canada. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2019\/file\/d202ed5bcfa858c15a9f383c3e386ab2-Paper.pdf"},{"key":"e_1_3_4_10_1","doi-asserted-by":"publisher","DOI":"10.1137\/1.9781611974997"},{"key":"e_1_3_4_11_1","doi-asserted-by":"publisher","DOI":"10.1017\/CBO9781139042918"},{"key":"e_1_3_4_12_1","doi-asserted-by":"publisher","DOI":"10.1145\/3320060"},{"issue":"10","key":"e_1_3_4_13_1","first-page":"281","article-title":"Random search for hyper-parameter optimization","volume":"13","author":"Bergstra J.","year":"2012","unstructured":"Bergstra, J., & Bengio, Y. (2012). Random search for hyper-parameter optimization. Journal of Machine Learning Research, 13(10), 281\u2013305. http:\/\/jmlr.org\/papers\/v13\/bergstra12a.html","journal-title":"Journal of Machine Learning Research"},{"key":"e_1_3_4_14_1","doi-asserted-by":"publisher","DOI":"10.48550\/arXiv.1810.05291"},{"key":"e_1_3_4_15_1","volume-title":"Parallel and distributed computation: Numerical methods","author":"Bertsekas D. P.","year":"1989","unstructured":"Bertsekas, D. P., & Tsitsiklis, J. N. (1989). Parallel and distributed computation: Numerical methods (Vol. 23). Prentice Hall Englewood Cliffs, NJ."},{"key":"e_1_3_4_16_1","first-page":"119","volume-title":"Advances in Neural Information Processing Systems","volume":"30","author":"Blanchard P.","year":"2017","unstructured":"Blanchard, P., El Mhamdi, E. M., Guerraoui, R., & Stainer, J. (2017). Machine learning with adversaries: Byzantine tolerant gradient descent. In: I. Guyon, U. Von Luxburg, S. Bengio, H. Wallach, R. Fergus, S. Vishwanathan, & R. Garnett (Eds.). Advances in Neural Information Processing Systems (Vol. 30. pp. 119\u2013129). Long Beach, CA, USA: Curran Associates, Inc. 31st Conference on Neural Information Processing Systems (NIPS 2017) https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2017\/file\/f4b9ec30ad9f68f89b29639786cb62ef-Paper.pdf"},{"issue":"9","key":"e_1_3_4_17_1","first-page":"142","article-title":"Online learning and stochastic approximations","volume":"17","author":"Bottou L.","year":"1998","unstructured":"Bottou, L. (1998). Online learning and stochastic approximations. Online Learning in Neural Networks, 17(9), 142.","journal-title":"Online Learning in Neural Networks"},{"key":"e_1_3_4_18_1","doi-asserted-by":"publisher","DOI":"10.1137\/16M1080173"},{"key":"e_1_3_4_19_1","first-page":"2004","article-title":"Subgradient methods","author":"Boyd S.","year":"2003","unstructured":"Boyd, S., Xiao, L., & Mutapcic, A. (2003). Subgradient methods. Lecture Notes of EE392o, Stanford University, Autumn Quarter, 2004\u20132005. https:\/\/see.stanford.edu\/materials\/lsocoee364b\/02-subgrad_method_notes.pdf","journal-title":"Lecture Notes of EE392o, Stanford University, Autumn Quarter"},{"key":"e_1_3_4_20_1","doi-asserted-by":"publisher","DOI":"10.1111\/1467-9884.00117"},{"key":"e_1_3_4_21_1","doi-asserted-by":"publisher","DOI":"10.1561\/2200000050"},{"key":"e_1_3_4_22_1","doi-asserted-by":"publisher","DOI":"10.1109\/GlobalSIP45357.2019.8969103"},{"key":"e_1_3_4_23_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIT.2005.858979"},{"key":"e_1_3_4_24_1","doi-asserted-by":"publisher","DOI":"10.1109\/TSP.2019.2946020"},{"key":"e_1_3_4_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICNN.1993.298667"},{"key":"e_1_3_4_26_1","doi-asserted-by":"publisher","DOI":"10.1109\/TNN.2009.2015974"},{"key":"e_1_3_4_27_1","doi-asserted-by":"publisher","DOI":"10.1145\/3055399.3055491"},{"key":"e_1_3_4_28_1","unstructured":"Chen J. Monga R. Bengio S. & Jozefowicz R. (2016). Revisiting distributed synchronous SGD. CoRR Abs\/1604.00981. https:\/\/arxiv.org\/pdf\/1604.00981"},{"key":"e_1_3_4_29_1","first-page":"903","volume-title":"Proceedings of the 35th International Conference on Machine Learning (ICML 2018)","volume":"80","author":"Chen L.","year":"2018","unstructured":"Chen, L., Wang, H., Charles, Z., & Papailiopoulos, D. (2018). Draco: Byzantine-resilient distributed training via redundant gradients. In: J. Dy & A. Krause (Eds.). Proceedings of the 35th International Conference on Machine Learning (ICML 2018) (Vol. 80. pp. 903\u2013912). PMLR. Stockholm, Sweden. 10\u201315 July 2018. https:\/\/proceedings.mlr.press\/v80\/chen18l.html"},{"key":"e_1_3_4_30_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11227-022-04733-8"},{"key":"e_1_3_4_31_1","doi-asserted-by":"publisher","DOI":"10.1109\/BigData.2018.8622598"},{"key":"e_1_3_4_32_1","doi-asserted-by":"publisher","DOI":"10.1145\/3154503"},{"key":"e_1_3_4_33_1","first-page":"571","volume-title":"11th USENIX Symposium on Operating Systems Design and Implementation (OSDI 14)","author":"Chilimbi T.","year":"2014","unstructured":"Chilimbi, T., Suzue, Y., Apacible, J., & Kalyanaraman, K. (2014). Project Adam: Building an efficient and scalable deep learning training system. In: 11th USENIX Symposium on Operating Systems Design and Implementation (OSDI 14) (pp. 571\u2013582). USENIX Association. Broomfield, CO. https:\/\/www.usenix.org\/conference\/osdi14\/technical-sessions\/presentation\/chilimbi"},{"key":"e_1_3_4_34_1","doi-asserted-by":"publisher","DOI":"10.1145\/2897518.2897647"},{"key":"e_1_3_4_35_1","unstructured":"Corporation N. (2021 December). GPUDirect enhancing data movement and access for GPUs. https:\/\/developer.nvidia.com\/gpudirect"},{"key":"e_1_3_4_36_1","volume-title":"Distributed systems: Concepts and design","author":"Coulouris G. F.","year":"2005","unstructured":"Coulouris, G. F., Dollimore, J., & Kindberg, T. (2005). Distributed systems: Concepts and design. Pearson Education."},{"key":"e_1_3_4_37_1","first-page":"81","volume-title":"Proceedings of Machine Learning and Systems (SysML 2019)","volume":"1","author":"Damaskinos G.","year":"2019","unstructured":"Damaskinos, G., El Mhamdi, E. M., Guerraoui, R., Guirguis, A. H. A., & Rouault, S. L. A. (2019). Aggregathor: Byzantine machine learning via robust gradient aggregation. In: A. Talwalkar, V. Smith, & M. Zaharia (Eds.). Proceedings of Machine Learning and Systems (SysML 2019) (Vol. 1. pp. 81\u2013106). Stanford, California. https:\/\/proceedings.mlsys.org\/paper_files\/paper\/2019\/file\/2f9b1b6b29361118f630783c19891ea0-Paper.pdf"},{"key":"e_1_3_4_38_1","first-page":"1145","volume-title":"Proceedings of the 35th International Conference on Machine Learning (ICML 2018)","volume":"80","author":"Damaskinos G.","year":"2018","unstructured":"Damaskinos, G., El Mhamdi, E. M., Guerraoui, R., Patra, R., & Taziki, M. (2018). Asynchronous byzantine machine learning (the case of SGD). In: J. Dy & A. Krause (Eds.). Proceedings of the 35th International Conference on Machine Learning (ICML 2018) (Vol. 80. pp. 1145\u20131154). PMLR. Stockholm, Sweden. 10\u201315 July 2018. https:\/\/proceedings.mlr.press\/v80\/damaskinos18a.html"},{"key":"e_1_3_4_39_1","unstructured":"Das D. Avancha S. Mudigere D. Vaidynathan K. Sridharan S. Kalamkar D. Kaul B. & Dubey P. (2016). Distributed deep learning using synchronous stochastic gradient descent. CoRR Abs\/1602.06709. http:\/\/arxiv.org\/abs\/1602.06709"},{"key":"e_1_3_4_40_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIT.2020.3035868"},{"key":"e_1_3_4_41_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.procs.2019.11.017"},{"key":"e_1_3_4_42_1","first-page":"1223","volume-title":"Advances in neural information processing systems (NeurIPS 2012)","author":"Dean J.","year":"2012","unstructured":"Dean, J., Corrado, G., Monga, R., Chen, K., Devin, M., Mao, M., Ranzato, M. A., Senior, A., Tucker, P., Yang, K., Le, Q., & Ng, A. (2012). Large scale distributed deep networks. In F. Pereira, C. J. Burges, L. Bottou, & K. Q. Weinberger (Eds.), Advances in neural information processing systems (NeurIPS 2012) (Vol. 25, pp. 1223\u20131231). Curran Associates, Inc. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2012\/file\/6aca97005c68f1206823815f66102863-Paper.pdf"},{"key":"e_1_3_4_43_1","first-page":"1646","volume-title":"Advances in neural information processing systems (NeurIPS 2014)","author":"Defazio A.","year":"2014","unstructured":"Defazio, A., Bach, F., & Lacoste-Julien, S. (2014). SAGA: A fast incremental gradient method with support for non-strongly convex composite objectives. In Z. Ghahramani, M. Welling, C. Cortes, N. Lawrence, & K. Q. Weinberger (Eds.), Advances in neural information processing systems (NeurIPS 2014) (Vol. 27, pp. 1646\u20131654). Curran Associates, Inc. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2014\/file\/ede7e2b6d13a41ddf9f4bdef84fdc737-Paper.pdf"},{"key":"e_1_3_4_44_1","doi-asserted-by":"publisher","DOI":"10.1137\/17M1126680"},{"key":"e_1_3_4_45_1","first-page":"999","volume-title":"Proceedings of the 34th International Conference on Machine Learning (ICML 2017)","volume":"70","author":"Diakonikolas I.","year":"2017","unstructured":"Diakonikolas, I., Kamath, G., Kane, D. M., Li, J., Moitra, A., & Stewart, A. (2017). Being robust (in high dimensions) can be practical. In D. Precup & Y. W. Teh (Eds.). Proceedings of the 34th International Conference on Machine Learning (ICML 2017) (Vol. 70. pp. 999\u20131008). PMLR. Sydney, Australia. 06\u201311 August 2017. https:\/\/proceedings.mlr.press\/v70\/diakonikolas17a.html"},{"key":"e_1_3_4_46_1","doi-asserted-by":"publisher","DOI":"10.1109\/CDC40024.2019.9029491"},{"key":"e_1_3_4_47_1","first-page":"1","volume-title":"Proceedings of the ICLR 2016 Workshop Track","author":"Dozat T.","year":"2016","unstructured":"Dozat, T. (2016). Incorporating Nesterov momentum into Adam. In Proceedings of the ICLR 2016 Workshop Track (pp. 1\u20134), San Juan, Puerto Rico. https:\/\/openreview.net\/pdf?id=OM0jvwB8jIp57ZJjtNEZ"},{"issue":"61","key":"e_1_3_4_48_1","first-page":"2121","article-title":"Adaptive subgradient methods for online learning and stochastic optimization","volume":"12","author":"Duchi J.","year":"2011","unstructured":"Duchi, J., Hazan, E., & Singer, Y. (2011). Adaptive subgradient methods for online learning and stochastic optimization. Journal of Machine Learning Research, 12(61), 2121\u20132159. http:\/\/jmlr.org\/papers\/v12\/duchi11a.html","journal-title":"Journal of Machine Learning Research"},{"key":"e_1_3_4_49_1","first-page":"2100","volume-title":"Advances in neural information processing systems (NeurIPS 2016)","author":"Dutta S.","year":"2016","unstructured":"Dutta, S., Cadambe, V., & Grover, P. (2016). Short-dot: Computing large linear transforms distributedly using coded short dot products. In D. Lee, M. Sugiyama, U. Luxburg, I. Guyon, & R. Garnett (Eds.), Advances in neural information processing systems (NeurIPS 2016) (Vol. 29, pp. 2100\u20132108). Curran Associates, Inc. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2016\/file\/aace49c7d80767cffec0e513ae886df0-Paper.pdf"},{"key":"e_1_3_4_50_1","doi-asserted-by":"publisher","DOI":"10.5075\/epfl-thesis-7218"},{"key":"e_1_3_4_51_1","first-page":"25044","volume-title":"Advances in neural information processing systems (NeurIPS 2021)","author":"El-Mhamdi E.-M.","year":"2021","unstructured":"El-Mhamdi, E.-M., Farhadkhani, S., Guerraoui, R., Guirguis, A., Hoang, L. N., & Rouault, S. (2021). Collaborative learning in the jungle (decentralized, byzantine, heterogeneous, asynchronous, and nonconvex learning). In M. Ranzato, A. Beygelzimer, Y. Dauphin, P. S. Liang, & J. W. Vaughan (Eds.), Advances in neural information processing systems (NeurIPS 2021) (Vol. 34, pp. 25044\u201325057). Curran Associates, Inc. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2021\/file\/d2cd33e9c0236a8c2d8bd3fa91ad3acf-Paper.pdf"},{"key":"e_1_3_4_52_1","doi-asserted-by":"publisher","DOI":"10.1145\/3382734.3405695"},{"key":"e_1_3_4_53_1","unstructured":"El-Mhamdi E.-M. Guerraoui R. & Rouault S. (2020). Distributed momentum for Byzantine-resilient learning. CoRR Abs\/2003.00010. Retrieved from. https:\/\/arxiv.org\/abs\/2003.00010"},{"key":"e_1_3_4_54_1","first-page":"3521","volume-title":"Proceedings of the 35th International Conference on Machine Learning (ICML 2018)","volume":"80","author":"El Mhamdi E. M.","year":"2018","unstructured":"El Mhamdi, E. M., Guerraoui, R., & Rouault, S. L. A. (2018). The hidden vulnerability of distributed learning in byzantium. In J. Dy & A. Krause (Eds.) Proceedings of the 35th International Conference on Machine Learning (ICML 2018) (Vol. 80. pp. 3521\u20133530). PMLR. Stockholm, Sweden. 10\u201315 July 2018. https:\/\/proceedings.mlr.press\/v80\/mhamdi18a.html"},{"key":"e_1_3_4_55_1","first-page":"89","volume-title":"Mathematics of stochastic manufacturing Systems, lectures in applied mathematics","author":"Ensor K. B.","year":"1997","unstructured":"Ensor, K. B., & Glynn, P. W. (1997). Stochastic optimization via grid search. In G. G. Yin & Q. Zhang (Eds.), Mathematics of stochastic manufacturing Systems, lectures in applied mathematics (Vol. 33, pp. 89\u2013100). American Mathematical Society, Providence, Rhode Island. https:\/\/web.stanford.edu\/~glynn\/papers\/1997\/EnsorG97.pdf"},{"key":"e_1_3_4_56_1","doi-asserted-by":"publisher","DOI":"10.1109\/TSIPN.2022.3188456"},{"key":"e_1_3_4_57_1","doi-asserted-by":"publisher","DOI":"10.1002\/nav.3800030109"},{"key":"e_1_3_4_58_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.cose.2023.103200"},{"key":"e_1_3_4_59_1","doi-asserted-by":"publisher","DOI":"10.1006\/jcss.1997.1504"},{"key":"e_1_3_4_60_1","doi-asserted-by":"publisher","DOI":"10.1016\/0898-1221(76)90003-1"},{"key":"e_1_3_4_61_1","doi-asserted-by":"publisher","DOI":"10.1109\/TCSVT.2021.3116976"},{"key":"e_1_3_4_62_1","first-page":"85","volume-title":"Advances in neural information processing systems","author":"Guo Y.","year":"2020","unstructured":"Guo, Y., Li, Q., & Chen, H. (2020). Backpropagating linearly improves transferability of adversarial examples. In H. Larochelle, M. Ranzato, R. Had Sell, M. F. Balcan, & H. Lin (Eds.), Advances in neural information processing systems (Vol. 33, pp. 85\u201395). Curran Associates, Inc. https:\/\/proceedings.neurips.cc\/paper\/2020\/file\/00e26af6ac3b1c1c49d7c3d79c60d000-Paper.pdf"},{"key":"e_1_3_4_63_1","unstructured":"Gupta N. & Vaidya N. H. (2019a). Byzantine fault tolerant distributed linear regression. CoRR abs\/1903.08752. http:\/\/arxiv.org\/abs\/1903.08752"},{"key":"e_1_3_4_64_1","doi-asserted-by":"publisher","DOI":"10.1109\/ALLERTON.2019.8919735"},{"key":"e_1_3_4_65_1","unstructured":"Gupta N. & Vaidya N. H. (2019c). Randomized reactive redundancy for byzantine fault-tolerance in parallelized learning. CoRR abs\/1912.09528. http:\/\/arxiv.org\/abs\/1912.09528"},{"key":"e_1_3_4_66_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11277-022-09624-y"},{"key":"e_1_3_4_67_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDM.2016.0028"},{"key":"e_1_3_4_68_1","unstructured":"Harris M. (2020 December). Inside Pascal: NVIDIA\u2019s newest computing platform. https:\/\/developer.nvidia.com\/blog\/inside-pascal\/"},{"key":"e_1_3_4_69_1","first-page":"267","volume-title":"Revised Selected Papers","volume":"13835","author":"Hillston J.","year":"2023","unstructured":"Hillston, J. (2023). A stochastic programming approach for an enhanced performance of a multi-committees byzantine fault tolerant algorithm. Revised Selected Papers. Euro-Par 2022: Parallel Processing Workshops: Euro-Par 2022 International Workshops (Vol. 13835. pp. 267, Glasgow, UK. August 22\u201326, 2022"},{"key":"e_1_3_4_70_1","doi-asserted-by":"publisher","DOI":"10.1145\/3133956.3134012"},{"key":"e_1_3_4_71_1","first-page":"1223","volume-title":"Advances in neural information processing systems (NeurIPS 2013)","author":"Ho Q.","year":"2013","unstructured":"Ho, Q., Cipar, J., Cui, H., Lee, S., Kim, J. K., Gibbons, P. B., Gibson, G. A., Ganger, G., & Xing, E. P. (2013). More effective distributed ML via a stale synchronous parallel parameter server. In C. J. Burges, L. Bottou, M. Welling, Z. Ghahramani, & K. Q. Weinberger (Eds.), Advances in neural information processing systems (NeurIPS 2013) (Vol. 26, pp. 1223\u20131231). Curran Associates, Inc. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2013\/file\/b7bb35b9c6ca2aee2df08cf09d7016c2-Paper.pdf"},{"key":"e_1_3_4_72_1","doi-asserted-by":"publisher","DOI":"10.1109\/MCOM.001.2000410"},{"key":"e_1_3_4_73_1","first-page":"629","volume-title":"14th USENIX Symposium on Networked Systems Design and Implementation (NSDI 17)","author":"Hsieh K.","year":"2017","unstructured":"Hsieh, K., Harlap, A., Vijaykumar, N., Konomis, D., Ganger, G. R., Gibbons, P. B., & Mutlu, O. (2017). Gaia: Geo-distributed machine learning approaching LAN speeds. In: 14th USENIX Symposium on Networked Systems Design and Implementation (NSDI 17) (pp. 629\u2013647). USENIX Association. Boston, MA. https:\/\/www.usenix.org\/conference\/nsdi17\/technical-sessions\/presentation\/hsieh"},{"key":"e_1_3_4_74_1","doi-asserted-by":"publisher","DOI":"10.1109\/MNET.129.2200346"},{"key":"e_1_3_4_75_1","unstructured":"Ibta) I. A. (2019 December). Infiniband - a low-latency high-bandwidth interconnect. https:\/\/www.infinibandta.org\/about-infiniband\/"},{"key":"e_1_3_4_76_1","unstructured":"Iqbal H. (2018 November). PlotNeuralnet: LaTeX code for making neural networks diagrams. https:\/\/github.com\/HarisIqbal88\/PlotNeuralNet"},{"key":"e_1_3_4_77_1","volume-title":"Fault tolerance in distributed systems","author":"Jalote P.","year":"1994","unstructured":"Jalote, P. (1994). Fault tolerance in distributed systems. Prentice-Hall, Inc."},{"key":"e_1_3_4_78_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2023.3264011"},{"key":"e_1_3_4_79_1","first-page":"5904","volume-title":"Advances in neural information processing systems (NeurIPS 2017)","author":"Jiang Z.","year":"2017","unstructured":"Jiang, Z., Balu, A., Hegde, C., & Sarkar, S. (2017). Collaborative deep learning in fixed topology networks. In I. Guyon, U. Von Luxburg, S. Bengio, H. Wallach, R. Fergus, S. Vishwanathan, & R. Garnett (Eds.), Advances in neural information processing systems (NeurIPS 2017) (Vol. 30, pp. 5904\u20135914). Curran Associates, Inc. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2017\/file\/a74c3bae3e13616104c1b25f9da1f11f-Paper.pdf"},{"key":"e_1_3_4_80_1","first-page":"1724","volume-title":"Proceedings of the 34th International Conference on Machine Learning (ICML 2017)","volume":"70","author":"Jin C.","year":"2017","unstructured":"Jin, C., Ge, R., Netrapalli, P., Kakade, S. M., & Jordan, M. I. (2017). How to escape saddle points efficiently. In: D. Precup & Y. W. Teh (Eds.). Proceedings of the 34th International Conference on Machine Learning (ICML 2017) (Vol. 70. pp. 1724\u20131732). PMLR. Sydney, Australia. 06\u201311 August 2017. https:\/\/proceedings.mlr.press\/v70\/jin17a.html"},{"key":"e_1_3_4_81_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICC.2019.8761674"},{"key":"e_1_3_4_82_1","unstructured":"Jin R. Huang Y. He X. Dai H. & Wu T. (2020). Stochastic-sign SGD for federated learning with theoretical guarantees. CoRR abs\/2002.10940. https:\/\/arxiv.org\/abs\/2002.10940"},{"key":"e_1_3_4_83_1","first-page":"315","volume-title":"Advances in neural information processing systems (NeurIPS 2013)","author":"Johnson R.","year":"2013","unstructured":"Johnson, R., & Zhang, T. (2013). Accelerating stochastic gradient descent using predictive variance reduction. In C. J. Burges, L. Bottou, M. Welling, Z. Ghahramani, & K. Q. Weinberger (Eds.), Advances in neural information processing systems (NeurIPS 2013) (Vol. 26, pp. 315\u2013323). Curran Associates, Inc. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2013\/file\/ac1dd209cbcc5e5d1c6e28598e8cbbe8-Paper.pdf"},{"key":"e_1_3_4_84_1","first-page":"4911","volume-title":"Proceedings of the 37th International Conference on Machine Learning (ICML 2020)","volume":"119","author":"Johnson T.","year":"2020","unstructured":"Johnson, T., Agrawal, P., Gu, H., & Guestrin, C. (2020). AdaScale SGD: A user-friendly algorithm for distributed training. In: H. Daum\u00e9 III & A. Singh (Eds.). Proceedings of the 37th International Conference on Machine Learning (ICML 2020) (Vol. 119. pp. 4911\u20134920). PMLR. 13\u201318 July 2020. https:\/\/proceedings.mlr.press\/v119\/johnson20a.html"},{"key":"e_1_3_4_85_1","doi-asserted-by":"publisher","DOI":"10.1561\/2200000083"},{"key":"e_1_3_4_86_1","first-page":"5434","volume-title":"Advances in neural information processing systems (NeurIPS 2017)","author":"Karakus C.","year":"2017","unstructured":"Karakus, C., Sun, Y., Diggavi, S., & Yin, W. (2017). Straggler mitigation in distributed optimization through data encoding. In I. Guyon, U. Von Luxburg, S. Bengio, H. Wallach, R. Fergus, S. Vishwanathan, & R. Garnett (Eds.), Advances in neural information processing systems (NeurIPS 2017) (Vol. 30, pp. 5434\u20135442). Curran Associates, Inc. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2017\/file\/663772ea088360f95bac3dc7ffb841be-Paper.pdf"},{"key":"e_1_3_4_87_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-1-4842-2766-4_8"},{"key":"e_1_3_4_88_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.procs.2022.12.227"},{"key":"e_1_3_4_89_1","doi-asserted-by":"publisher","DOI":"10.48550\/arXiv.1412.6980"},{"key":"e_1_3_4_90_1","first-page":"2698","volume-title":"Proceedings of the 35th International Conference on Machine Learning (ICML 2018)","volume":"80","author":"Kleinberg B.","year":"2018","unstructured":"Kleinberg, B., Li, Y., & Yuan, Y. (2018). An alternative view: When does SGD escape local minima? In: J. Dy & A. Krause (Eds.). Proceedings of the 35th International Conference on Machine Learning (ICML 2018) (Vol. 80. pp. 2698\u20132707). PMLR. Stockholm, Sweden. https:\/\/proceedings.mlr.press\/v80\/kleinberg18a.html"},{"key":"e_1_3_4_91_1","doi-asserted-by":"publisher","DOI":"10.1109\/SYNASC.2016.041"},{"key":"e_1_3_4_92_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.jksuci.2018.09.021"},{"key":"e_1_3_4_93_1","doi-asserted-by":"publisher","DOI":"10.1145\/357172.357176"},{"key":"e_1_3_4_94_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIT.2017.2736066"},{"key":"e_1_3_4_95_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-91704-7_5"},{"key":"e_1_3_4_96_1","unstructured":"Le Roux N. Schmidt M. & Bach F.) Advances in Neural Information Processing Systems (NeurIPS 2012). (2012). A stochastic gradient method with an exponential convergence rate for finite training sets. In F. Pereira C. J. Burges L. Bottou & K. Q. Weinberger (Eds.). (Vol. 25 pp. 2663\u20132671). Curran Associates Inc. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2012\/file\/905056c1ac1dad141560467e0a99e1cf-Paper.pdf"},{"key":"e_1_3_4_97_1","doi-asserted-by":"publisher","DOI":"10.1109\/TNN.2007.912320"},{"key":"e_1_3_4_98_1","doi-asserted-by":"publisher","DOI":"10.1007\/s00500-021-06496-5"},{"key":"e_1_3_4_99_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v33i01.33011544"},{"key":"e_1_3_4_100_1","doi-asserted-by":"publisher","DOI":"10.1145\/2640087.2644155"},{"key":"e_1_3_4_101_1","first-page":"19","volume-title":"Advances in neural information processing systems (NeurIPS 2014)","author":"Li M.","year":"2014","unstructured":"Li, M., Andersen, D. G., Smola, A. J., & Yu, K. (2014b). Communication efficient distributed machine learning with the parameter server. In Z. Ghahramani, M. Welling, C. Cortes, N. Lawrence, & K. Q. Weinberger (Eds.), Advances in neural information processing systems (NeurIPS 2014) (Vol. 27, pp. 19\u201327). Curran Associates, Inc. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2014\/file\/1ff1de774005f8da13f42943881c655f-Paper.pdf"},{"key":"e_1_3_4_102_1","first-page":"2","volume-title":"Big Learning NIPS Workshop","volume":"6","author":"Li M.","year":"2013","unstructured":"Li, M., Zhou, L., Yang, Z., Li, A., Xia, F., Andersen, D. G., & Smola, A. (2013). Parameter server for distributed machine learning. In: Big Learning NIPS Workshop (Vol. 6. pp. 2). Lake Tahoe, Nevada. http:\/\/www.cs.cmu.edu\/afs\/cs\/Web\/People\/muli\/file\/ps.pdf"},{"key":"e_1_3_4_103_1","first-page":"5330","volume-title":"Advances in neural information processing systems (NeurIPS 2017)","author":"Lian X.","year":"2017","unstructured":"Lian, X., Zhang, C., Zhang, H., Hsieh, C.-J., Zhang, W., & Liu, J. (2017). Can decentralized algorithms outperform centralized algorithms? A case study for decentralized parallel stochastic gradient descent. In I. Guyon, U. Von Luxburg, S. Bengio, H. Wallach, R. Fergus, S. Vishwanathan, & R. Garnett (Eds.), Advances in neural information processing systems (NeurIPS 2017) (Vol. 30, pp. 5330\u20135340). Curran Associates, Inc. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2017\/file\/f75526659f31040afeb61cb7133e4e6d-Paper.pdf"},{"key":"e_1_3_4_104_1","first-page":"3043","volume-title":"Proceedings of the 35th International Conference on Machine Learning (ICML 2018)","volume":"80","author":"Lian X.","year":"2018","unstructured":"Lian, X., Zhang, W., Zhang, C., & Liu, J. (2018). Asynchronous decentralized parallel stochastic gradient descent. In: J. Dy & A. Krause (Eds.). Proceedings of the 35th International Conference on Machine Learning (ICML 2018) (Vol. 80. pp. 3043\u20133052). Stockholm, Sweden, PMLR. https:\/\/proceedings.mlr.press\/v80\/lian18a.html"},{"key":"e_1_3_4_105_1","volume-title":"International Conference on Learning Representations (ICLR 2020), Addis Ababa","author":"Liu C.","year":"2020","unstructured":"Liu, C., & Belkin, M. (2020). Accelerating SGD with momentum for over-parameterized learning. In: International Conference on Learning Representations (ICLR 2020), Addis Ababa, Ethiopia. https:\/\/openreview.net\/pdf?id=r1gixp4FPH"},{"key":"e_1_3_4_106_1","unstructured":"Liu S. (2021). A survey on fault-tolerance in distributed optimization and machine learning. CoRR Abs\/2106.08545. https:\/\/arxiv.org\/abs\/2106.08545"},{"key":"e_1_3_4_107_1","doi-asserted-by":"publisher","DOI":"10.1145\/3477114.3488761"},{"key":"e_1_3_4_108_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2019.2959220"},{"key":"e_1_3_4_109_1","volume-title":"Distributed algorithms","author":"Lynch N. A.","year":"1996","unstructured":"Lynch, N. A. (1996). Distributed algorithms. Elsevier."},{"key":"e_1_3_4_110_1","first-page":"2113","volume-title":"Proceedings of the 32nd International Conference on Machine Learning","volume":"37","author":"Maclaurin D.","year":"2015","unstructured":"Maclaurin, D., Duvenaud, D., & Adams, R. (2015). Gradient-based hyperparameter optimization through reversible learning. In: F. Bach & D. Blei (Eds.). Proceedings of the 32nd International Conference on Machine Learning (Vol. 37. pp. 2113\u20132122). Lille, France, PMLR. https:\/\/proceedings.mlr.press\/v37\/maclaurin15.html"},{"key":"e_1_3_4_111_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICIEM59379.2023.10166343"},{"key":"e_1_3_4_112_1","first-page":"1273","volume-title":"Proceedings of the 20th International Conference on Artificial Intelligence and Statistics (AISTATS) 2017, Fort Lauderdale","volume":"54","author":"McMahan B.","year":"2017","unstructured":"McMahan, B., Moore, E., Ramage, D., Hampson, S., & Arcas, B. A. Y. (2017). Communication-efficient learning of deep networks from decentralized data. In: A. Singh & J. Zhu (Eds.). Proceedings of the 20th International Conference on Artificial Intelligence and Statistics (AISTATS) 2017, Fort Lauderdale (Vol. 54. pp. 1273\u20131282). PMLR, Florida, USA. https:\/\/proceedings.mlr.press\/v54\/mcmahan17a.html"},{"key":"e_1_3_4_113_1","doi-asserted-by":"publisher","DOI":"10.1109\/SP.2019.00029"},{"key":"e_1_3_4_114_1","doi-asserted-by":"publisher","DOI":"10.1049\/iet-wss.2018.5229"},{"key":"e_1_3_4_115_1","doi-asserted-by":"publisher","DOI":"10.1049\/iet-wss.2019.0106"},{"key":"e_1_3_4_116_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.tcs.2017.03.018"},{"key":"e_1_3_4_117_1","article-title":"Bitcoin: A peer-to-peer electronic cash system","volume":"21260","author":"Nakamoto S.","year":"2008","unstructured":"Nakamoto, S. (2008a). Bitcoin: A peer-to-peer electronic cash system. Decentralized Business Review, 21260. https:\/\/static.upbitcare.com\/931b8bfc-f0e0-4588-be6e-b98a27991df1.pdf","journal-title":"Decentralized Business Review"},{"key":"e_1_3_4_118_1","unstructured":"Nakamoto S. (2008b). A peer-to-peer electronic cash system. Bitcoin. https:\/\/bitcoin.org\/bitcoin.pdf"},{"key":"e_1_3_4_119_1","doi-asserted-by":"publisher","DOI":"10.1109\/SP.2019.00065"},{"key":"e_1_3_4_120_1","doi-asserted-by":"publisher","DOI":"10.1109\/TAC.2008.2009515"},{"issue":"3","key":"e_1_3_4_121_1","first-page":"543","article-title":"A method for solving the convex programming problem with convergence rate o (1\/k2)","volume":"269","author":"Nesterov Y. E.","year":"1983","unstructured":"Nesterov, Y. E. (1983). A method for solving the convex programming problem with convergence rate o (1\/k2). Doklady Akademii nauk SSSR, 269(3), 543\u2013547. https:\/\/www.mathnet.ru\/eng\/dan\/v269\/i3\/p543","journal-title":"Doklady Akademii nauk SSSR"},{"key":"e_1_3_4_122_1","volume-title":"The 2nd international workshop on federated learning for data privacy and confidentiality","author":"Orekondy T.","year":"2019","unstructured":"Orekondy, T., Oh, S. J., Zhang, Y., Schiele, B., & Fritz, M. (2019). Gradient-Leaks: Understanding deanonymization in federated learning. In: The 2nd international workshop on federated learning for data privacy and confidentiality. Canada. https:\/\/hdl.handle.net\/21.11116\/0000-0005-8E47-C"},{"key":"e_1_3_4_123_1","doi-asserted-by":"publisher","DOI":"10.1145\/3597062.3597280"},{"key":"e_1_3_4_124_1","doi-asserted-by":"publisher","DOI":"10.5121\/ijsc.2011.2204"},{"key":"e_1_3_4_125_1","doi-asserted-by":"publisher","DOI":"10.1145\/322186.322188"},{"key":"e_1_3_4_126_1","doi-asserted-by":"publisher","DOI":"10.1016\/0041-5553(64)90137-5"},{"key":"e_1_3_4_127_1","doi-asserted-by":"publisher","DOI":"10.1016\/S0893-6080(98)00116-6"},{"key":"e_1_3_4_128_1","first-page":"5220","volume-title":"Proceedings of the 36th International Conference on Machine Learning","volume":"97","author":"Qiao A.","year":"2019","unstructured":"Qiao, A., Aragam, B., Zhang, B., & Xing, E. (2019). Fault tolerance in iterative-convergent machine learning. In: K. Chaudhuri & R. Salakhutdinov (Eds.). Proceedings of the 36th International Conference on Machine Learning (Vol. 97. pp. 5220\u20135230). PMLR, Long Beach, California. https:\/\/proceedings.mlr.press\/v97\/qiao19a.html"},{"key":"e_1_3_4_129_1","unstructured":"Raicu I. Foster I. Szalay A. & Turcu G. (2006). Astroportal: A science gateway for large-scale astronomy data analysis. TeraGrid \u201906: Advancing scientific discovery Indianapolis Indiana. (12\u201315). http:\/\/datasys.cs.iit.edu\/reports\/2006_AstroPortal_2-page-handout_v2_high.pdf"},{"key":"e_1_3_4_130_1","first-page":"10320","volume-title":"Advances in neural information processing systems","author":"Rajput S.","year":"2019","unstructured":"Rajput, S., Wang, H., Charles, Z., & Papailiopoulos, D. (2019). DETOX: A redundancy-based framework for faster and more robust gradient aggregation. In H. Wallach, H. Larochelle, A. Beygelzimer, F. d\u2019Alch\u00e9-Buc, E. Fox, & R. Garnett (Eds.), Advances in neural information processing systems (Vol. 32, pp. 10320\u201310330). Curran Associates, Inc. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2019\/file\/415185ea244ea2b2bedeb0449b926802-Paper.pdf"},{"key":"e_1_3_4_131_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00130"},{"key":"e_1_3_4_132_1","doi-asserted-by":"publisher","DOI":"10.3390\/su11143974"},{"key":"e_1_3_4_133_1","first-page":"4305","volume-title":"Proceedings of the 35th International Conference on Machine Learning, Stockholm, Sweden","volume":"80","author":"Raviv N.","year":"2018","unstructured":"Raviv, N., Tandon, R., Dimakis, A., & Tamo, I. (2018). Gradient coding from cyclic MDS codes and expander graphs. In: J. Dy & A. Krause (Eds.). Proceedings of the 35th International Conference on Machine Learning, Stockholm, Sweden (Vol. 80. pp. 4305\u20134313). PMLR. https:\/\/proceedings.mlr.press\/v80\/raviv18a.html"},{"key":"e_1_3_4_134_1","doi-asserted-by":"publisher","DOI":"10.2200\/S00294ED1V01Y201009DCT003"},{"key":"e_1_3_4_135_1","volume-title":"Concurrent programming: Algorithms, principles, and foundations","author":"Raynal M.","year":"2012","unstructured":"Raynal, M. (2012). Concurrent programming: Algorithms, principles, and foundations. Springer Science & Business Media."},{"key":"e_1_3_4_136_1","first-page":"2647","volume-title":"Advances in neural information processing systems","author":"Reddi S. J.","year":"2015","unstructured":"Reddi, S. J., Hefny, A., Sra, S., Poczos, B., & Smola, A. J. (2015). On variance reduction in stochastic gradient descent and its asynchronous variants. In C. Cortes, N. Lawrence, D. Lee, M. Sugiyama, & R. Garnett (Eds.), Advances in neural information processing systems (Vol. 28, pp. 2647\u20132655). Curran Associates, Inc. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2015\/file\/d010396ca8abf6ead8cacc2c2f2f26c7-Paper.pdf"},{"key":"e_1_3_4_137_1","doi-asserted-by":"publisher","DOI":"10.1214\/aoms\/1177729586"},{"key":"e_1_3_4_138_1","unstructured":"Ruder S. (2016). An overview of gradient descent optimization algorithms. CoRR Abs\/1609.04747. http:\/\/arxiv.org\/abs\/1609.04747"},{"key":"e_1_3_4_139_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2022.3144217"},{"key":"e_1_3_4_140_1","doi-asserted-by":"publisher","DOI":"10.1002\/widm.1249"},{"key":"e_1_3_4_141_1","doi-asserted-by":"publisher","DOI":"10.1145\/190.357399"},{"key":"e_1_3_4_142_1","doi-asserted-by":"publisher","DOI":"10.1109\/JPROC.2015.2494218"},{"key":"e_1_3_4_143_1","doi-asserted-by":"publisher","DOI":"10.1017\/CBO9781107298019"},{"key":"e_1_3_4_144_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.micpro.2023.104796"},{"key":"e_1_3_4_145_1","doi-asserted-by":"publisher","DOI":"10.1002\/cpe.7610"},{"key":"e_1_3_4_146_1","doi-asserted-by":"publisher","DOI":"10.1109\/DASC\/PiCom\/DataCom\/CyberSciTec.2018.000-4"},{"key":"e_1_3_4_147_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDCS.2019.00220"},{"key":"e_1_3_4_148_1","doi-asserted-by":"publisher","DOI":"10.1109\/72.471356"},{"key":"e_1_3_4_149_1","doi-asserted-by":"publisher","DOI":"10.1145\/3133956.3134077"},{"key":"e_1_3_4_150_1","first-page":"2377","volume-title":"Advances in Neural Information Processing Systems","volume":"28","author":"Srivastava R. K.","year":"2015","unstructured":"Srivastava, R. K., Greff, K., & Schmidhuber, J. (2015). Training very deep networks. In: C. Cortes, N. Lawrence, D. Lee, M. Sugiyama, & R. Garnett (Eds.). Advances in Neural Information Processing Systems (Vol. 28. pp. 2377\u20132385). Curran Associates, Inc. 28th Conference on Neural Information Processing Systems (NIPS 2015), Montreal, Canada. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2015\/file\/215a71a12769b056c3c32e7299f1c5ed-Paper.pdf"},{"key":"e_1_3_4_151_1","unstructured":"Statista. (2023 July). Amount of data created consumed and stored 2010\u20132025. https:\/\/www.statista.com\/statistics\/871513\/worldwide-data-created\/"},{"key":"e_1_3_4_152_1","doi-asserted-by":"publisher","DOI":"10.4230\/LIPIcs.ITCS.2018.45"},{"key":"e_1_3_4_153_1","doi-asserted-by":"publisher","DOI":"10.48550\/arXiv.1511.01821"},{"key":"e_1_3_4_154_1","doi-asserted-by":"publisher","DOI":"10.1145\/2933057.2933105"},{"key":"e_1_3_4_155_1","doi-asserted-by":"publisher","DOI":"10.1145\/3322205.3311083"},{"key":"e_1_3_4_156_1","doi-asserted-by":"publisher","DOI":"10.1109\/TCYB.2019.2950779"},{"key":"e_1_3_4_157_1","volume-title":"Paper presented at the 2nd International Conference on Learning Representations, ICLR 2014","author":"Szegedy C.","year":"2014","unstructured":"Szegedy, C., Zaremba, W., Sutskever, I., Bruna, J., Erhan, D., Goodfellow, I., & Fergus, R. (2014). Intriguing properties of neural networks. Paper presented at the 2nd International Conference on Learning Representations, ICLR 2014, Banff, Canada."},{"key":"e_1_3_4_158_1","first-page":"3368","volume-title":"Proceedings of the 34th International Conference on Machine Learning","volume":"70","author":"Tandon R.","year":"2017","unstructured":"Tandon, R., Lei, Q., Dimakis, A. G., & Karampatziakis, N. (2017). Gradient coding: Avoiding stragglers in distributed learning. In: D. Precup & Y. W. Teh (Eds.). Proceedings of the 34th International Conference on Machine Learning (Vol. 70. pp. 3368\u20133376). PMLR, Sydney, Australia. https:\/\/proceedings.mlr.press\/v70\/tandon17a.html"},{"key":"e_1_3_4_159_1","doi-asserted-by":"publisher","DOI":"10.1109\/ITAIC.2019.8785560"},{"issue":"2","key":"e_1_3_4_160_1","first-page":"26","article-title":"Lecture 6.5-RMSProp: Divide the gradient by a running average of its recent magnitude","volume":"4","author":"Tieleman T.","year":"2012","unstructured":"Tieleman, T., & Hinton, G. (2012). Lecture 6.5-RMSProp: Divide the gradient by a running average of its recent magnitude. COURSERA: Neural Networks for Machine Learning, 4(2), 26\u201331.","journal-title":"COURSERA: Neural Networks for Machine Learning"},{"key":"e_1_3_4_161_1","doi-asserted-by":"publisher","DOI":"10.1145\/3377454"},{"key":"e_1_3_4_162_1","doi-asserted-by":"publisher","DOI":"10.23919\/CCC55666.2022.9902233"},{"key":"e_1_3_4_163_1","doi-asserted-by":"publisher","DOI":"10.48550\/arXiv.1612.00334"},{"key":"e_1_3_4_164_1","doi-asserted-by":"publisher","DOI":"10.1109\/TDSC.2019.2952332"},{"key":"e_1_3_4_165_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10107-015-0892-3"},{"key":"e_1_3_4_166_1","doi-asserted-by":"publisher","DOI":"10.1109\/BigData52589.2021.9671583"},{"key":"e_1_3_4_167_1","doi-asserted-by":"publisher","DOI":"10.24963\/ijcai.2019\/670"},{"key":"e_1_3_4_168_1","doi-asserted-by":"publisher","DOI":"10.48550\/arXiv.1802.10116"},{"key":"e_1_3_4_169_1","doi-asserted-by":"publisher","DOI":"10.48550\/arXiv.1805.09682"},{"key":"e_1_3_4_170_1","first-page":"6893","volume-title":"Proceedings of the 36th International Conference on Machine Learning","volume":"97","author":"Xie C.","year":"2019","unstructured":"Xie, C., Koyejo, S., & Gupta, I. (2019). Zeno: Distributed stochastic gradient descent with suspicion-based fault-tolerance. In: K. Chaudhuri & R. Salakhutdinov (Eds.). Proceedings of the 36th International Conference on Machine Learning (Vol. 97. pp. 6893\u20136901). PMLR, Long Beach, California. https:\/\/proceedings.mlr.press\/v97\/xie19b.html"},{"key":"e_1_3_4_171_1","first-page":"10495","volume-title":"Proceedings of the 37th International Conference on Machine Learning","volume":"119","author":"Xie C.","year":"2020","unstructured":"Xie, C., Koyejo, S., & Gupta, I. (2020). Zeno++: Robust Fully Asynchronous SGD. In: H. Daum\u00e9 III & A. Singh (Eds.). Proceedings of the 37th International Conference on Machine Learning (Vol. 119. pp. 10495\u201310503). PMLR. Virtual. https:\/\/proceedings.mlr.press\/v119\/xie20c.html"},{"key":"e_1_3_4_172_1","doi-asserted-by":"publisher","DOI":"10.1016\/J.ENG.2016.02.008"},{"key":"e_1_3_4_173_1","doi-asserted-by":"publisher","DOI":"10.1109\/TC.2022.3169436"},{"key":"e_1_3_4_174_1","doi-asserted-by":"publisher","DOI":"10.1109\/CDC40024.2019.9029245"},{"key":"e_1_3_4_175_1","first-page":"11751","volume-title":"Proceedings of the 38th International Conference on Machine Learning","volume":"139","author":"Yang Y.-R.","year":"2021","unstructured":"Yang, Y.-R., & Li, W.-J. (2021). BASGD: Buffered asynchronous SGD for byzantine learning. In: M. Meila & T. Zhang (Eds.). Proceedings of the 38th International Conference on Machine Learning (Vol. 139. pp. 11751\u201311761). PMLR, Virtual. http:\/\/proceedings.mlr.press\/v139\/yang21e\/yang21e.pdf"},{"key":"e_1_3_4_176_1","doi-asserted-by":"publisher","DOI":"10.1109\/TSIPN.2019.2928176"},{"key":"e_1_3_4_177_1","doi-asserted-by":"publisher","DOI":"10.1109\/MSP.2020.2973345"},{"key":"e_1_3_4_178_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICPR.2018.8546189"},{"key":"e_1_3_4_179_1","first-page":"5650","volume-title":"Proceedings of the 35th International Conference on Machine Learning, Stockholm, Sweden","volume":"80","author":"Yin D.","year":"2018","unstructured":"Yin, D., Chen, Y., Kannan, R., & Bartlett, P. (2018). Byzantine-robust distributed learning: Towards optimal statistical rates. In: J. Dy & A. Krause (Eds.). Proceedings of the 35th International Conference on Machine Learning, Stockholm, Sweden (Vol. 80. pp. 5650\u20135659). PMLR. http:\/\/proceedings.mlr.press\/v80\/yin18a\/yin18a.pdf"},{"key":"e_1_3_4_180_1","first-page":"7074","volume-title":"Proceedings of the 36th International Conference on Machine Learning, Long Beach, California","volume":"97","author":"Yin D.","year":"2019","unstructured":"Yin, D., Chen, Y., Kannan, R., & Bartlett, P. (2019). Defending against saddle point attack in Byzantine-robust distributed learning. In: K. Chaudhuri & R. Salakhutdinov (Eds.). Proceedings of the 36th International Conference on Machine Learning, Long Beach, California (Vol. 97. pp. 7074\u20137084). PMLR. http:\/\/proceedings.mlr.press\/v97\/yin19a\/yin19a.pdf"},{"key":"e_1_3_4_181_1","doi-asserted-by":"publisher","DOI":"10.1109\/TNNLS.2018.2886017"},{"key":"e_1_3_4_182_1","first-page":"7","volume-title":"8th USENIX symposium on operating Systems design and implementation (OSDI\u201908)","author":"Zaharia M.","year":"2008","unstructured":"Zaharia, M., Konwinski, A., Joseph, A. D., Katz, R. H., & Stoica, I. (2008, December 2008, 8\u201310). Improving MapReduce performance in heterogeneous environments. In: 8th USENIX symposium on operating Systems design and implementation (OSDI\u201908) (Vol. 8, p. 7). USENIX Association. https:\/\/www.usenix.org\/legacy\/event\/osdi08\/tech\/full_papers\/zaharia\/zaharia.pdf"},{"key":"e_1_3_4_183_1","doi-asserted-by":"publisher","DOI":"10.48550\/arXiv.1212.5701"},{"key":"e_1_3_4_184_1","doi-asserted-by":"publisher","DOI":"10.1145\/3636553"},{"key":"e_1_3_4_185_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2018.2820899"},{"key":"e_1_3_4_186_1","first-page":"685","volume-title":"Advances in neural information processing systems","author":"Zhang S.","year":"2015","unstructured":"Zhang, S., Choromanska, A. E., & LeCun, Y. (2015). Deep learning with elastic averaging SGD. In C. Cortes, N. Lawrence, D. Lee, M. Sugiyama, & R. Garnett (Eds.), Advances in neural information processing systems (Vol. 28, pp. 685\u2013693). Curran Associates, Inc. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2015\/file\/d18f655c3fce66ca401d5f38b48c89af-Paper.pdf"},{"key":"e_1_3_4_187_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2013.6638950"},{"key":"e_1_3_4_188_1","doi-asserted-by":"publisher","DOI":"10.1109\/CSCloud-EdgeCom58631.2023.00080"},{"key":"e_1_3_4_189_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDCS.2019.00150"},{"key":"e_1_3_4_190_1","doi-asserted-by":"publisher","DOI":"10.1109\/JIOT.2020.3017377"},{"key":"e_1_3_4_191_1","doi-asserted-by":"publisher","DOI":"10.1109\/BigDataCongress.2017.85"},{"key":"e_1_3_4_192_1","doi-asserted-by":"publisher","DOI":"10.1007\/s00521-003-0353-4"}],"container-title":["Journal of Experimental &amp; Theoretical Artificial Intelligence"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.tandfonline.com\/doi\/pdf\/10.1080\/0952813X.2024.2391778","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,11,2]],"date-time":"2025-11-02T22:42:33Z","timestamp":1762123353000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.tandfonline.com\/doi\/full\/10.1080\/0952813X.2024.2391778"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,9,12]]},"references-count":191,"journal-issue":{"issue":"8","published-print":{"date-parts":[[2025,11,17]]}},"alternative-id":["10.1080\/0952813X.2024.2391778"],"URL":"https:\/\/doi.org\/10.1080\/0952813x.2024.2391778","relation":{},"ISSN":["0952-813X","1362-3079"],"issn-type":[{"value":"0952-813X","type":"print"},{"value":"1362-3079","type":"electronic"}],"subject":[],"published":{"date-parts":[[2024,9,12]]},"assertion":[{"value":"The publishing and review policy for this title is described in its Aims & Scope.","order":1,"name":"peerreview_statement","label":"Peer Review Statement"},{"value":"http:\/\/www.tandfonline.com\/action\/journalInformation?show=aimsScope&journalCode=teta20","URL":"http:\/\/www.tandfonline.com\/action\/journalInformation?show=aimsScope&journalCode=teta20","order":2,"name":"aims_and_scope_url","label":"Aim & Scope"},{"value":"2023-05-06","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2024-07-16","order":2,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2024-09-12","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}