{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,10]],"date-time":"2026-07-10T16:41:01Z","timestamp":1783701661893,"version":"3.55.0"},"publisher-location":"New York, NY, USA","reference-count":49,"publisher":"ACM","license":[{"start":{"date-parts":[[2021,6,21]],"date-time":"2021-06-21T00:00:00Z","timestamp":1624233600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"DOI":"10.13039\/100000001","name":"NSF (National Science Foundation)","doi-asserted-by":"publisher","award":["CNS-1717763, CCF-1618776"],"award-info":[{"award-number":["CNS-1717763, CCF-1618776"]}],"id":[{"id":"10.13039\/100000001","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2021,6,21]]},"DOI":"10.1145\/3431379.3460650","type":"proceedings-article","created":{"date-parts":[[2021,6,17]],"date-time":"2021-06-17T04:09:26Z","timestamp":1623902966000},"page":"189-200","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":23,"title":["Q-adaptive"],"prefix":"10.1145","author":[{"given":"Yao","family":"Kang","sequence":"first","affiliation":[{"name":"Illinois Institute of Technology, Chicago, IL, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Xin","family":"Wang","sequence":"additional","affiliation":[{"name":"Illinois Institute of Technology, Chicago, IL, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Zhiling","family":"Lan","sequence":"additional","affiliation":[{"name":"Illinois Institute of Technology, Chicago, IL, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2021,6,21]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1145\/2184319.2184335"},{"key":"e_1_3_2_1_2_1","volume-title":"Cray XC series network","author":"Alverson Bob","year":"2012","unstructured":"Bob Alverson , Edwin Froese , Larry Kaplan , and Duncan Roweth . 2012. Cray XC series network . Cray Inc., White Paper WP-Aries 01-1112 ( 2012 ). Bob Alverson, Edwin Froese, Larry Kaplan, and Duncan Roweth. 2012. Cray XC series network. Cray Inc., White Paper WP-Aries01-1112 (2012)."},{"key":"e_1_3_2_1_3_1","volume-title":"Packet routing in dynamically changing networks: A reinforcement learning approach. Advances in neural information processing systems","author":"Boyan Justin","year":"1993","unstructured":"Justin Boyan and Michael Littman . 1993. Packet routing in dynamically changing networks: A reinforcement learning approach. Advances in neural information processing systems , Vol. 6 ( 1993 ), 671--678. Justin Boyan and Michael Littman. 1993. Packet routing in dynamically changing networks: A reinforcement learning approach. Advances in neural information processing systems, Vol. 6 (1993), 671--678."},{"key":"e_1_3_2_1_4_1","volume-title":"Predictive Q-routing: A memory-based reinforcement learning approach to adaptive traffic control. In Advances in Neural Information Processing Systems. 945--951.","author":"Choi Samuel PM","year":"1996","unstructured":"Samuel PM Choi and Dit-Yan Yeung . 1996 . Predictive Q-routing: A memory-based reinforcement learning approach to adaptive traffic control. In Advances in Neural Information Processing Systems. 945--951. Samuel PM Choi and Dit-Yan Yeung. 1996. Predictive Q-routing: A memory-based reinforcement learning approach to adaptive traffic control. In Advances in Neural Information Processing Systems. 945--951."},{"key":"e_1_3_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1145\/3126908.3126926"},{"key":"e_1_3_2_1_6_1","first-page":"746","article-title":"The dynamics of reinforcement learning in cooperative multiagent systems","volume":"1998","author":"Claus Caroline","year":"1998","unstructured":"Caroline Claus and Craig Boutilier . 1998 . The dynamics of reinforcement learning in cooperative multiagent systems . AAAI\/IAAI , Vol. 1998 , 746 -- 752 (1998), 2. Caroline Claus and Craig Boutilier. 1998. The dynamics of reinforcement learning in cooperative multiagent systems. AAAI\/IAAI, Vol. 1998, 746--752 (1998), 2.","journal-title":"AAAI\/IAAI"},{"key":"e_1_3_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/71.127260"},{"key":"e_1_3_2_1_8_1","unstructured":"William J Dally and Charles L Seitz. 1988. Deadlock-free message routing in multiprocessor interconnection networks. (1988).  William J Dally and Charles L Seitz. 1988. Deadlock-free message routing in multiprocessor interconnection networks. (1988)."},{"key":"e_1_3_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1145\/3295500.3356196"},{"key":"e_1_3_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/SC.2012.39"},{"key":"e_1_3_2_1_11_1","volume-title":"Deep Reinforcement Agent for Scheduling in HPC. In 2021 IEEE International Parallel and Distributed Processing Symposium (IPDPS).","author":"Fan Yuping","year":"2021","unstructured":"Yuping Fan , Taylor Childers , Paul Rich , William Allcock , Michael Papka , and Zhiling Lan . 2021 . Deep Reinforcement Agent for Scheduling in HPC. In 2021 IEEE International Parallel and Distributed Processing Symposium (IPDPS). Yuping Fan, Taylor Childers, Paul Rich, William Allcock, Michael Papka, and Zhiling Lan. 2021. Deep Reinforcement Agent for Scheduling in HPC. In 2021 IEEE International Parallel and Distributed Processing Symposium (IPDPS)."},{"key":"e_1_3_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-92040-5_15"},{"key":"e_1_3_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICPP.2012.46"},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1109\/SC.2014.33"},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1145\/1555754.1555783"},{"key":"e_1_3_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1145\/3316480.3325517"},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISCA.2008.19"},{"key":"e_1_3_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1103\/PhysRevB.47.558"},{"key":"e_1_3_2_1_19_1","volume-title":"In Proceedings of the Seventeenth International Conference on Machine Learning. Citeseer.","author":"Lauer Martin","year":"2000","unstructured":"Martin Lauer and Martin Riedmiller . 2000 . An algorithm for distributed reinforcement learning in cooperative multi-agent systems . In In Proceedings of the Seventeenth International Conference on Machine Learning. Citeseer. Martin Lauer and Martin Riedmiller. 2000. An algorithm for distributed reinforcement learning in cooperative multi-agent systems. In In Proceedings of the Seventeenth International Conference on Machine Learning. Citeseer."},{"key":"e_1_3_2_1_20_1","volume-title":"Continuous control with deep reinforcement learning. arXiv preprint arXiv:1509.02971","author":"Lillicrap Timothy P","year":"2015","unstructured":"Timothy P Lillicrap , Jonathan J Hunt , Alexander Pritzel , Nicolas Heess , Tom Erez , Yuval Tassa , David Silver , and Daan Wierstra . 2015. Continuous control with deep reinforcement learning. arXiv preprint arXiv:1509.02971 ( 2015 ). Timothy P Lillicrap, Jonathan J Hunt, Alexander Pritzel, Nicolas Heess, Tom Erez, Yuval Tassa, David Silver, and Daan Wierstra. 2015. Continuous control with deep reinforcement learning. arXiv preprint arXiv:1509.02971 (2015)."},{"key":"e_1_3_2_1_21_1","volume-title":"Zili Meng, and Mohammad Alizadeh.","author":"Mao Hongzi","year":"2019","unstructured":"Hongzi Mao , Malte Schwarzkopf , Shaileshh Bojja Venkatakrishnan , Zili Meng, and Mohammad Alizadeh. 2019 . Learning scheduling algorithms for data processing clusters. In Proceedings of the ACM Special Interest Group on Data Communication. 270--288. Hongzi Mao, Malte Schwarzkopf, Shaileshh Bojja Venkatakrishnan, Zili Meng, and Mohammad Alizadeh. 2019. Learning scheduling algorithms for data processing clusters. In Proceedings of the ACM Special Interest Group on Data Communication. 270--288."},{"key":"e_1_3_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1109\/IROS.2007.4399095"},{"key":"e_1_3_2_1_23_1","volume-title":"Device placement optimization with reinforcement learning. arXiv preprint arXiv:1706.04972","author":"Mirhoseini Azalia","year":"2017","unstructured":"Azalia Mirhoseini , Hieu Pham , Quoc V Le , Benoit Steiner , Rasmus Larsen , Yuefeng Zhou , Naveen Kumar , Mohammad Norouzi , Samy Bengio , and Jeff Dean . 2017. Device placement optimization with reinforcement learning. arXiv preprint arXiv:1706.04972 ( 2017 ). Azalia Mirhoseini, Hieu Pham, Quoc V Le, Benoit Steiner, Rasmus Larsen, Yuefeng Zhou, Naveen Kumar, Mohammad Norouzi, Samy Bengio, and Jeff Dean. 2017. Device placement optimization with reinforcement learning. arXiv preprint arXiv:1706.04972 (2017)."},{"key":"e_1_3_2_1_24_1","volume-title":"et almbox","author":"Mnih Volodymyr","year":"2015","unstructured":"Volodymyr Mnih , Koray Kavukcuoglu , David Silver , Andrei A Rusu , Joel Veness , Marc G Bellemare , Alex Graves , Martin Riedmiller , Andreas K Fidjeland , Georg Ostrovski , et almbox . 2015 . Human-level control through deep reinforcement learning. nature, Vol. 518 , 7540 (2015), 529--533. Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et almbox. 2015. Human-level control through deep reinforcement learning. nature, Vol. 518, 7540 (2015), 529--533."},{"key":"e_1_3_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-20656-7_1"},{"key":"e_1_3_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICSMC.1998.726708"},{"key":"e_1_3_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/IJCNN.2002.1007796"},{"key":"e_1_3_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1109\/SC.2002.10019"},{"key":"e_1_3_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.sysconle.2016.02.020"},{"key":"e_1_3_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1145\/3295500.3356208"},{"key":"e_1_3_2_1_31_1","volume-title":"Deep Neural Networks for Network Routing. In 2019 International Joint Conference on Neural Networks (IJCNN). IEEE, 1--8.","author":"Reis Joao","year":"2019","unstructured":"Joao Reis , Miguel Rocha , Truong Khoa Phan , David Griffin , Franck Le , and Miguel Rio . 2019 . Deep Neural Networks for Network Routing. In 2019 International Joint Conference on Neural Networks (IJCNN). IEEE, 1--8. Joao Reis, Miguel Rocha, Truong Khoa Phan, David Griffin, Franck Le, and Miguel Rio. 2019. Deep Neural Networks for Network Routing. In 2019 International Joint Conference on Neural Networks (IJCNN). IEEE, 1--8."},{"key":"e_1_3_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1145\/1964218.1964225"},{"key":"e_1_3_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.1145\/1150019.1136488"},{"key":"e_1_3_2_1_34_1","volume-title":"An In-Depth Analysis of the Slingshot Interconnect. In 2020 SC20: International Conference for High Performance Computing, Networking, Storage and Analysis (SC). IEEE Computer Society, 481--494","author":"Sensi Daniele","year":"2020","unstructured":"Daniele Sensi , Salvatore Girolamo , Kim McMahon , Duncan Roweth , and Torsten Hoefler . 2020 . An In-Depth Analysis of the Slingshot Interconnect. In 2020 SC20: International Conference for High Performance Computing, Networking, Storage and Analysis (SC). IEEE Computer Society, 481--494 . Daniele Sensi, Salvatore Girolamo, Kim McMahon, Duncan Roweth, and Torsten Hoefler. 2020. An In-Depth Analysis of the Slingshot Interconnect. In 2020 SC20: International Conference for High Performance Computing, Networking, Storage and Analysis (SC). IEEE Computer Society, 481--494."},{"key":"e_1_3_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/HiPINEB.2017.11"},{"key":"e_1_3_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1063\/1.874055"},{"key":"e_1_3_2_1_37_1","volume-title":"Reinforcement learning: An introduction","author":"Sutton Richard S","unstructured":"Richard S Sutton and Andrew G Barto . 2018. Reinforcement learning: An introduction . MIT press . Richard S Sutton and Andrew G Barto. 2018. Reinforcement learning: An introduction. MIT press."},{"key":"e_1_3_2_1_38_1","volume-title":"Computational electrodynamics: the finite-difference time-domain method","author":"Taflove Allen","unstructured":"Allen Taflove and Susan C Hagness . 2005. Computational electrodynamics: the finite-difference time-domain method . Artech house. Allen Taflove and Susan C Hagness. 2005. Computational electrodynamics: the finite-difference time-domain method. Artech house."},{"key":"e_1_3_2_1_39_1","volume-title":"In: Proc. of the 18th Int. Conf. on Machine Learning. Citeseer.","author":"Tao Nigel","year":"2001","unstructured":"Nigel Tao , Jonathan Baxter , and Lex Weaver . 2001 . A multi-agent, policy-gradient approach to network routing . In In: Proc. of the 18th Int. Conf. on Machine Learning. Citeseer. Nigel Tao, Jonathan Baxter, and Lex Weaver. 2001. A multi-agent, policy-gradient approach to network routing. In In: Proc. of the 18th Int. Conf. on Machine Learning. Citeseer."},{"key":"e_1_3_2_1_40_1","unstructured":"top500.org. 2020. Top500 list. https:\/\/www.top500.org\/lists\/top500\/2020\/11\/  top500.org. 2020. Top500 list. https:\/\/www.top500.org\/lists\/top500\/2020\/11\/"},{"key":"e_1_3_2_1_41_1","doi-asserted-by":"publisher","DOI":"10.1145\/1995896.1995932"},{"key":"e_1_3_2_1_42_1","doi-asserted-by":"publisher","DOI":"10.1145\/3152434.3152441"},{"key":"e_1_3_2_1_43_1","volume-title":"Learning to design circuits. arXiv preprint arXiv:1812.02734","author":"Wang Hanrui","year":"2018","unstructured":"Hanrui Wang , Jiacheng Yang , Hae-Seung Lee , and Song Han . 2018. Learning to design circuits. arXiv preprint arXiv:1812.02734 ( 2018 ). Hanrui Wang, Jiacheng Yang, Hae-Seung Lee, and Song Han. 2018. Learning to design circuits. arXiv preprint arXiv:1812.02734 (2018)."},{"key":"e_1_3_2_1_44_1","volume-title":"Union: An Automatic Workload Manager for Accelerating Network Simulation. In 2020 IEEE International Parallel and Distributed Processing Symposium (IPDPS). IEEE, 821--830","author":"Wang Xin","year":"2020","unstructured":"Xin Wang , Misbah Mubarak , Yao Kang , Robert B Ross , and Zhiling Lan . 2020 . Union: An Automatic Workload Manager for Accelerating Network Simulation. In 2020 IEEE International Parallel and Distributed Processing Symposium (IPDPS). IEEE, 821--830 . Xin Wang, Misbah Mubarak, Yao Kang, Robert B Ross, and Zhiling Lan. 2020. Union: An Automatic Workload Manager for Accelerating Network Simulation. In 2020 IEEE International Parallel and Distributed Processing Symposium (IPDPS). IEEE, 821--830."},{"key":"e_1_3_2_1_45_1","doi-asserted-by":"publisher","DOI":"10.1109\/CLUSTER49012.2020.00021"},{"key":"e_1_3_2_1_46_1","doi-asserted-by":"publisher","DOI":"10.1109\/HPCA.2015.7056051"},{"key":"e_1_3_2_1_47_1","doi-asserted-by":"publisher","DOI":"10.5555\/3014904.3014990"},{"key":"e_1_3_2_1_48_1","volume-title":"Experiences with ML-Driven Design: A NoC Case Study. In 2020 IEEE International Symposium on High Performance Computer Architecture (HPCA). IEEE, 637--648","author":"Yin Jieming","year":"2020","unstructured":"Jieming Yin , Subhash Sethumurugan , Yasuko Eckert , Chintan Patel , Alan Smith , Eric Morton , Mark Oskin , Natalie Enright Jerger , and Gabriel H Loh . 2020 . Experiences with ML-Driven Design: A NoC Case Study. In 2020 IEEE International Symposium on High Performance Computer Architecture (HPCA). IEEE, 637--648 . Jieming Yin, Subhash Sethumurugan, Yasuko Eckert, Chintan Patel, Alan Smith, Eric Morton, Mark Oskin, Natalie Enright Jerger, and Gabriel H Loh. 2020. Experiences with ML-Driven Design: A NoC Case Study. In 2020 IEEE International Symposium on High Performance Computer Architecture (HPCA). IEEE, 637--648."},{"key":"e_1_3_2_1_49_1","doi-asserted-by":"publisher","DOI":"10.1109\/TSMC.2020.3012832"}],"event":{"name":"HPDC '21: The 30th International Symposium on High-Performance Parallel and Distributed Computing","location":"Virtual Event Sweden","acronym":"HPDC '21","sponsor":["University of Arizona University of Arizona","SIGHPC ACM Special Interest Group on High Performance Computing, Special Interest Group on High Performance Computing","SIGARCH ACM Special Interest Group on Computer Architecture"]},"container-title":["Proceedings of the 30th International Symposium on High-Performance Parallel and Distributed Computing"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3431379.3460650","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/abs\/10.1145\/3431379.3460650","content-type":"text\/html","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3431379.3460650","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3431379.3460650","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T21:24:46Z","timestamp":1750195486000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3431379.3460650"}},"subtitle":["A Multi-Agent Reinforcement Learning Based Routing on Dragonfly Network"],"short-title":[],"issued":{"date-parts":[[2021,6,21]]},"references-count":49,"alternative-id":["10.1145\/3431379.3460650","10.1145\/3431379"],"URL":"https:\/\/doi.org\/10.1145\/3431379.3460650","relation":{},"subject":[],"published":{"date-parts":[[2021,6,21]]},"assertion":[{"value":"2021-06-21","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}