{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,25]],"date-time":"2026-06-25T14:50:57Z","timestamp":1782399057019,"version":"3.54.5"},"publisher-location":"New York, NY, USA","reference-count":36,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,10,17]],"date-time":"2022-10-17T00:00:00Z","timestamp":1665964800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"the National Natural Science Foundation of China","award":["11901578"],"award-info":[{"award-number":["11901578"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,10,17]]},"DOI":"10.1145\/3511808.3557292","type":"proceedings-article","created":{"date-parts":[[2022,10,16]],"date-time":"2022-10-16T01:29:57Z","timestamp":1665883797000},"page":"842-851","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":5,"title":["Diverse Effective Relationship Exploration for Cooperative Multi-Agent Reinforcement Learning"],"prefix":"10.1145","author":[{"given":"Hao","family":"Jiang","sequence":"first","affiliation":[{"name":"Academy of Military Science, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yuntao","family":"Liu","sequence":"additional","affiliation":[{"name":"Academy of Military Science, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Shengze","family":"Li","sequence":"additional","affiliation":[{"name":"Academy of Military Science, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Jieyuan","family":"Zhang","sequence":"additional","affiliation":[{"name":"Academy of Military Science, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Xinhai","family":"Xu","sequence":"additional","affiliation":[{"name":"Academy of Military Science, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Donghong","family":"Liu","sequence":"additional","affiliation":[{"name":"Academy of Military Science, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2022,10,17]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1145\/3319619.3321894"},{"key":"e_1_3_2_1_2_1","volume-title":"International Conference on Distributed Artificial Intelligence. Springer, 185--205","author":"Chen Liheng","year":"2021","unstructured":"Liheng Chen , Hongyi Guo , Yali Du , Fei Fang , Haifeng Zhang , Weinan Zhang , and Yong Yu . 2021 . Signal instructed coordination in cooperative multi-agent reinforcement learning . In International Conference on Distributed Artificial Intelligence. Springer, 185--205 . Liheng Chen, Hongyi Guo, Yali Du, Fei Fang, Haifeng Zhang, Weinan Zhang, and Yong Yu. 2021. Signal instructed coordination in cooperative multi-agent reinforcement learning. In International Conference on Distributed Artificial Intelligence. Springer, 185--205."},{"key":"e_1_3_2_1_3_1","volume-title":"Advances in Neural Information Processing Systems","volume":"34","author":"Chenghao Li","year":"2021","unstructured":"Li Chenghao , Tonghan Wang , Chengjie Wu , Qianchuan Zhao , Jun Yang , and Chongjie Zhang . 2021 . Celebrating diversity in shared multi-agent reinforcement learning . Advances in Neural Information Processing Systems , Vol. 34 (2021). Li Chenghao, Tonghan Wang, Chengjie Wu, Qianchuan Zhao, Jun Yang, and Chongjie Zhang. 2021. Celebrating diversity in shared multi-agent reinforcement learning. Advances in Neural Information Processing Systems, Vol. 34 (2021)."},{"key":"e_1_3_2_1_4_1","volume-title":"Nando De Freitas, and Shimon Whiteson.","author":"Foerster Jakob","year":"2016","unstructured":"Jakob Foerster , Ioannis Alexandros Assael , Nando De Freitas, and Shimon Whiteson. 2016 . Learning to communicate with deep multi-agent reinforcement learning. Advances in neural information processing systems, Vol. 29 (2016). Jakob Foerster, Ioannis Alexandros Assael, Nando De Freitas, and Shimon Whiteson. 2016. Learning to communicate with deep multi-agent reinforcement learning. Advances in neural information processing systems, Vol. 29 (2016)."},{"key":"e_1_3_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v32i1.11794"},{"key":"e_1_3_2_1_6_1","volume-title":"Filip De Turck, and Pieter Abbeel","author":"Houthooft Rein","year":"2016","unstructured":"Rein Houthooft , Xi Chen , Yan Duan , John Schulman , Filip De Turck, and Pieter Abbeel . 2016 . Curiosity-driven exploration in deep reinforcement learning via bayesian neural networks. (2016). Rein Houthooft, Xi Chen, Yan Duan, John Schulman, Filip De Turck, and Pieter Abbeel. 2016. Curiosity-driven exploration in deep reinforcement learning via bayesian neural networks. (2016)."},{"key":"e_1_3_2_1_7_1","volume-title":"Graph convolutional reinforcement learning. arXiv preprint arXiv:1810.09202","author":"Jiang Jiechuan","year":"2018","unstructured":"Jiechuan Jiang , Chen Dun , Tiejun Huang , and Zongqing Lu. 2018. Graph convolutional reinforcement learning. arXiv preprint arXiv:1810.09202 ( 2018 ). Jiechuan Jiang, Chen Dun, Tiejun Huang, and Zongqing Lu. 2018. Graph convolutional reinforcement learning. arXiv preprint arXiv:1810.09202 (2018)."},{"key":"e_1_3_2_1_8_1","volume-title":"Autonomous robot vehicles","author":"Khatib Oussama","unstructured":"Oussama Khatib . 1986. Real-time obstacle avoidance for manipulators and mobile robots . In Autonomous robot vehicles . Springer , 396--404. Oussama Khatib. 1986. Real-time obstacle avoidance for manipulators and mobile robots. In Autonomous robot vehicles. Springer, 396--404."},{"key":"e_1_3_2_1_9_1","volume-title":"A maximum mutual information framework for multi-agent reinforcement learning. arXiv preprint arXiv:2006.02732","author":"Kim Woojun","year":"2020","unstructured":"Woojun Kim , Whiyoung Jung , Myungsik Cho , and Youngchul Sung . 2020. A maximum mutual information framework for multi-agent reinforcement learning. arXiv preprint arXiv:2006.02732 ( 2020 ). Woojun Kim, Whiyoung Jung, Myungsik Cho, and Youngchul Sung. 2020. A maximum mutual information framework for multi-agent reinforcement learning. arXiv preprint arXiv:2006.02732 (2020)."},{"key":"e_1_3_2_1_10_1","volume-title":"Revisiting the master-slave architecture in multi-agent deep reinforcement learning. arXiv preprint arXiv:1712.07305","author":"Kong Xiangyu","year":"2017","unstructured":"Xiangyu Kong , Bo Xin , Fangchen Liu , and Yizhou Wang . 2017. Revisiting the master-slave architecture in multi-agent deep reinforcement learning. arXiv preprint arXiv:1712.07305 ( 2017 ). Xiangyu Kong, Bo Xin, Fangchen Liu, and Yizhou Wang. 2017. Revisiting the master-slave architecture in multi-agent deep reinforcement learning. arXiv preprint arXiv:1712.07305 (2017)."},{"key":"e_1_3_2_1_11_1","unstructured":"Karol Kurach Anton Raichuk Piotr Sta\u0144czyk Micha\u0142 Zajk\u0105c Olivier Bachem Lasse Espeholt Carlos Riquelme Damien Vincent Marcin Michalski Olivier Bousquet etal 2019. Google research football: A novel reinforcement learning environment. arXiv preprint arXiv:1907.11180 (2019).  Karol Kurach Anton Raichuk Piotr Sta\u0144czyk Micha\u0142 Zajk\u0105c Olivier Bachem Lasse Espeholt Carlos Riquelme Damien Vincent Marcin Michalski Olivier Bousquet et al. 2019. Google research football: A novel reinforcement learning environment. arXiv preprint arXiv:1907.11180 (2019)."},{"key":"e_1_3_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v33i01.33014213"},{"key":"e_1_3_2_1_13_1","volume-title":"OpenAI Pieter Abbeel, and Igor Mordatch","author":"Lowe Ryan","year":"2017","unstructured":"Ryan Lowe , Yi I Wu , Aviv Tamar , Jean Harb , OpenAI Pieter Abbeel, and Igor Mordatch . 2017 . Multi-agent actor-critic for mixed cooperative-competitive environments. In Advances in neural information processing systems. 6379--6390. Ryan Lowe, Yi I Wu, Aviv Tamar, Jean Harb, OpenAI Pieter Abbeel, and Igor Mordatch. 2017. Multi-agent actor-critic for mixed cooperative-competitive environments. In Advances in neural information processing systems. 6379--6390."},{"key":"e_1_3_2_1_14_1","volume-title":"A concise introduction to decentralized POMDPs","author":"Oliehoek Frans A","unstructured":"Frans A Oliehoek and Christopher Amato . 2016. A concise introduction to decentralized POMDPs . Springer . Frans A Oliehoek and Christopher Amato. 2016. A concise introduction to decentralized POMDPs. Springer."},{"key":"e_1_3_2_1_15_1","unstructured":"Sankar K Pal and Sushmita Mitra. 1992. Multilayer perceptron fuzzy sets classifiaction. (1992).  Sankar K Pal and Sushmita Mitra. 1992. Multilayer perceptron fuzzy sets classifiaction. (1992)."},{"key":"e_1_3_2_1_16_1","volume-title":"Multiagent bidirectionally-coordinated nets: Emergence of human-level coordination in learning to play starcraft combat games. arXiv preprint arXiv:1703.10069","author":"Peng Peng","year":"2017","unstructured":"Peng Peng , Ying Wen , Yaodong Yang , Quan Yuan , Zhenkun Tang , Haitao Long , and Jun Wang . 2017. Multiagent bidirectionally-coordinated nets: Emergence of human-level coordination in learning to play starcraft combat games. arXiv preprint arXiv:1703.10069 ( 2017 ). Peng Peng, Ying Wen, Yaodong Yang, Quan Yuan, Zhenkun Tang, Haitao Long, and Jun Wang. 2017. Multiagent bidirectionally-coordinated nets: Emergence of human-level coordination in learning to play starcraft combat games. arXiv preprint arXiv:1703.10069 (2017)."},{"key":"e_1_3_2_1_17_1","volume-title":"David Feil-Seifer, and Aria Nefian.","author":"Pham Huy Xuan","year":"2018","unstructured":"Huy Xuan Pham , Hung Manh La , David Feil-Seifer, and Aria Nefian. 2018 . Cooperative and distributed reinforcement learning of drones for field coverage. arXiv preprint arXiv:1803.07250 (2018). Huy Xuan Pham, Hung Manh La, David Feil-Seifer, and Aria Nefian. 2018. Cooperative and distributed reinforcement learning of drones for field coverage. arXiv preprint arXiv:1803.07250 (2018)."},{"key":"e_1_3_2_1_18_1","volume-title":"Weighted qmix: Expanding monotonic value function factorisation for deep multi-agent reinforcement learning. Advances in neural information processing systems","author":"Rashid Tabish","year":"2020","unstructured":"Tabish Rashid , Gregory Farquhar , Bei Peng , and Shimon Whiteson . 2020. Weighted qmix: Expanding monotonic value function factorisation for deep multi-agent reinforcement learning. Advances in neural information processing systems , Vol. 33 ( 2020 ), 10199--10210. Tabish Rashid, Gregory Farquhar, Bei Peng, and Shimon Whiteson. 2020. Weighted qmix: Expanding monotonic value function factorisation for deep multi-agent reinforcement learning. Advances in neural information processing systems, Vol. 33 (2020), 10199--10210."},{"key":"e_1_3_2_1_19_1","volume-title":"International Conference on Machine Learning. PMLR, 4295--4304","author":"Rashid Tabish","year":"2018","unstructured":"Tabish Rashid , Mikayel Samvelyan , Christian Schroeder , Gregory Farquhar , Jakob Foerster , and Shimon Whiteson . 2018 . Qmix: Monotonic value function factorisation for deep multi-agent reinforcement learning . In International Conference on Machine Learning. PMLR, 4295--4304 . Tabish Rashid, Mikayel Samvelyan, Christian Schroeder, Gregory Farquhar, Jakob Foerster, and Shimon Whiteson. 2018. Qmix: Monotonic value function factorisation for deep multi-agent reinforcement learning. In International Conference on Machine Learning. PMLR, 4295--4304."},{"key":"e_1_3_2_1_20_1","volume-title":"Gregory Farquhar, Nantas Nardelli, Tim GJ Rudner, Chia-Man Hung, Philip HS Torr, Jakob Foerster, and Shimon Whiteson.","author":"Samvelyan Mikayel","year":"2019","unstructured":"Mikayel Samvelyan , Tabish Rashid , Christian Schroeder De Witt , Gregory Farquhar, Nantas Nardelli, Tim GJ Rudner, Chia-Man Hung, Philip HS Torr, Jakob Foerster, and Shimon Whiteson. 2019 . The starcraft multi-agent challenge. arXiv preprint arXiv:1902.04043 (2019). Mikayel Samvelyan, Tabish Rashid, Christian Schroeder De Witt, Gregory Farquhar, Nantas Nardelli, Tim GJ Rudner, Chia-Man Hung, Philip HS Torr, Jakob Foerster, and Shimon Whiteson. 2019. The starcraft multi-agent challenge. arXiv preprint arXiv:1902.04043 (2019)."},{"key":"e_1_3_2_1_21_1","volume-title":"Dynamics-aware unsupervised discovery of skills. arXiv preprint arXiv:1907.01657","author":"Sharma Archit","year":"2019","unstructured":"Archit Sharma , Shixiang Gu , Sergey Levine , Vikash Kumar , and Karol Hausman . 2019. Dynamics-aware unsupervised discovery of skills. arXiv preprint arXiv:1907.01657 ( 2019 ). Archit Sharma, Shixiang Gu, Sergey Levine, Vikash Kumar, and Karol Hausman. 2019. Dynamics-aware unsupervised discovery of skills. arXiv preprint arXiv:1907.01657 (2019)."},{"key":"e_1_3_2_1_22_1","unstructured":"Arambam James Singh Akshat Kumar and Hoong Chuin Lau. 2020. Hierarchical multiagent reinforcement learning for maritime traffic management. (2020).  Arambam James Singh Akshat Kumar and Hoong Chuin Lau. 2020. Hierarchical multiagent reinforcement learning for maritime traffic management. (2020)."},{"key":"e_1_3_2_1_23_1","volume-title":"QTRAN: Improved Value Transformation for Cooperative Multi-Agent Reinforcement Learning. https:\/\/openreview.net\/forum?id=TlS3LBoDj3Z","author":"Son Kyunghwan","year":"2021","unstructured":"Kyunghwan Son , Sungsoo Ahn , Roben D. Delos Reyes , Jinwoo Shin , and Yung Yi . 2021 . QTRAN: Improved Value Transformation for Cooperative Multi-Agent Reinforcement Learning. https:\/\/openreview.net\/forum?id=TlS3LBoDj3Z Kyunghwan Son, Sungsoo Ahn, Roben D. Delos Reyes, Jinwoo Shin, and Yung Yi. 2021. QTRAN: Improved Value Transformation for Cooperative Multi-Agent Reinforcement Learning. https:\/\/openreview.net\/forum?id=TlS3LBoDj3Z"},{"key":"e_1_3_2_1_24_1","volume-title":"David Earl Hostallero, and Yung Yi.","author":"Son Kyunghwan","year":"2019","unstructured":"Kyunghwan Son , Daewoo Kim , Wan Ju Kang , David Earl Hostallero, and Yung Yi. 2019 . Qtran : Learning to factorize with transformation for cooperative multi-agent reinforcement learning. arXiv preprint arXiv:1905.05408 (2019). Kyunghwan Son, Daewoo Kim, Wan Ju Kang, David Earl Hostallero, and Yung Yi. 2019. Qtran: Learning to factorize with transformation for cooperative multi-agent reinforcement learning. arXiv preprint arXiv:1905.05408 (2019)."},{"key":"e_1_3_2_1_25_1","volume-title":"Multi type mean field reinforcement learning. arXiv preprint arXiv:2002.02513","author":"Subramanian Sriram Ganapathi","year":"2020","unstructured":"Sriram Ganapathi Subramanian , Pascal Poupart , Matthew E. Taylor , and Nidhi Hegde . 2020. Multi type mean field reinforcement learning. arXiv preprint arXiv:2002.02513 ( 2020 ). Sriram Ganapathi Subramanian, Pascal Poupart, Matthew E. Taylor, and Nidhi Hegde. 2020. Multi type mean field reinforcement learning. arXiv preprint arXiv:2002.02513 (2020)."},{"key":"e_1_3_2_1_26_1","unstructured":"Sainbayar Sukhbaatar Rob Fergus etal 2016. Learning multiagent communication with backpropagation. Advances in neural information processing systems Vol. 29 (2016).  Sainbayar Sukhbaatar Rob Fergus et al. 2016. Learning multiagent communication with backpropagation. Advances in neural information processing systems Vol. 29 (2016)."},{"key":"e_1_3_2_1_27_1","volume-title":"Vinicius Zambaldi, Max Jaderberg, Marc Lanctot, Nicolas Sonnerat, Joel Z Leibo, Karl Tuyls, et al.","author":"Sunehag Peter","year":"2017","unstructured":"Peter Sunehag , Guy Lever , Audrunas Gruslys , Wojciech Marian Czarnecki , Vinicius Zambaldi, Max Jaderberg, Marc Lanctot, Nicolas Sonnerat, Joel Z Leibo, Karl Tuyls, et al. 2017 . Value-decomposition networks for cooperative multi-agent learning. arXiv preprint arXiv:1706.05296 (2017). Peter Sunehag, Guy Lever, Audrunas Gruslys, Wojciech Marian Czarnecki, Vinicius Zambaldi, Max Jaderberg, Marc Lanctot, Nicolas Sonnerat, Joel Z Leibo, Karl Tuyls, et al. 2017. Value-decomposition networks for cooperative multi-agent learning. arXiv preprint arXiv:1706.05296 (2017)."},{"key":"e_1_3_2_1_28_1","volume-title":"Foundations and Trends\u00ae in Machine Learning","volume":"1","author":"Wainwright Martin J","year":"2008","unstructured":"Martin J Wainwright , Michael I Jordan , 2008 . Graphical models, exponential families, and variational inference . Foundations and Trends\u00ae in Machine Learning , Vol. 1 , 1--2 (2008), 1--305. Martin J Wainwright, Michael I Jordan, et al. 2008. Graphical models, exponential families, and variational inference. Foundations and Trends\u00ae in Machine Learning, Vol. 1, 1--2 (2008), 1--305."},{"key":"e_1_3_2_1_29_1","volume-title":"Qplex: Duplex dueling multi-agent q-learning. arXiv preprint arXiv:2008.01062","author":"Wang Jianhao","year":"2020","unstructured":"Jianhao Wang , Zhizhou Ren , Terry Liu , Yang Yu , and Chongjie Zhang . 2020 b. Qplex: Duplex dueling multi-agent q-learning. arXiv preprint arXiv:2008.01062 (2020). Jianhao Wang, Zhizhou Ren, Terry Liu, Yang Yu, and Chongjie Zhang. 2020b. Qplex: Duplex dueling multi-agent q-learning. arXiv preprint arXiv:2008.01062 (2020)."},{"key":"e_1_3_2_1_30_1","volume-title":"Rode: Learning roles to decompose multi-agent tasks. arXiv preprint arXiv:2010.01523","author":"Wang Tonghan","year":"2020","unstructured":"Tonghan Wang , Tarun Gupta , Anuj Mahajan , Bei Peng , Shimon Whiteson , and Chongjie Zhang . 2020 a. Rode: Learning roles to decompose multi-agent tasks. arXiv preprint arXiv:2010.01523 (2020). Tonghan Wang, Tarun Gupta, Anuj Mahajan, Bei Peng, Shimon Whiteson, and Chongjie Zhang. 2020a. Rode: Learning roles to decompose multi-agent tasks. arXiv preprint arXiv:2010.01523 (2020)."},{"key":"e_1_3_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.1109\/ROBOT.1989.100007"},{"key":"e_1_3_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v33i01.33011206"},{"key":"e_1_3_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCA.2018.8444355"},{"key":"e_1_3_2_1_34_1","unstructured":"Vinicius Zambaldi David Raposo Adam Santoro Victor Bapst Yujia Li Igor Babuschkin Karl Tuyls David Reichert Timothy Lillicrap Edward Lockhart etal 2018. Relational deep reinforcement learning. arXiv preprint arXiv:1806.01830 (2018).  Vinicius Zambaldi David Raposo Adam Santoro Victor Bapst Yujia Li Igor Babuschkin Karl Tuyls David Reichert Timothy Lillicrap Edward Lockhart et al. 2018. Relational deep reinforcement learning. arXiv preprint arXiv:1806.01830 (2018)."},{"key":"e_1_3_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.5555\/2900423.2900545"},{"key":"e_1_3_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v35i13.17357"}],"event":{"name":"CIKM '22: The 31st ACM International Conference on Information and Knowledge Management","location":"Atlanta GA USA","acronym":"CIKM '22","sponsor":["SIGWEB ACM Special Interest Group on Hypertext, Hypermedia, and Web","SIGIR ACM Special Interest Group on Information Retrieval"]},"container-title":["Proceedings of the 31st ACM International Conference on Information &amp; Knowledge Management"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3511808.3557292","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3511808.3557292","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T17:49:28Z","timestamp":1750182568000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3511808.3557292"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,10,17]]},"references-count":36,"alternative-id":["10.1145\/3511808.3557292","10.1145\/3511808"],"URL":"https:\/\/doi.org\/10.1145\/3511808.3557292","relation":{},"subject":[],"published":{"date-parts":[[2022,10,17]]},"assertion":[{"value":"2022-10-17","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}