{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2023,9,30]],"date-time":"2023-09-30T14:43:51Z","timestamp":1696085031508},"reference-count":0,"publisher":"IOS Press","isbn-type":[{"value":"9781643684369","type":"print"},{"value":"9781643684376","type":"electronic"}],"license":[{"start":{"date-parts":[[2023,9,28]],"date-time":"2023-09-28T00:00:00Z","timestamp":1695859200000},"content-version":"unspecified","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by-nc\/4.0\/"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2023,9,28]]},"abstract":"<jats:p>Cooperative multi-agent reinforcement learning (Co-MARL) commonly employs different parameter sharing mechanisms, such as full and partial sharing. However, imprudent application of these mechanisms can potentially constrain policy diversity and limit cooperation flexibility. Recent methods that group agents into distinct sharing categories often exhibit poor performance due to challenges in precisely differentiating agents and neglecting the issue of promoting cooperation among these categories. To address these issues, we introduce a dynamic selective parameter sharing mechanism embedded with multi-level reasoning abstractions (DSPS-MA). Our approach uses self-comparison sequences to infer agents\u2019 abstract concepts, defining the differences between agents and allowing them to dynamically select partners to share parameters based on these abstract concepts. We also design an intrinsic reward to offer comprehensive collaboration guidance for agents, and introduce a policy cosine similarity regularization term to ensure sufficient policy diversity. Empirical evaluations demonstrate that our approach yields higher returns and faster convergence than state-of-the-art methods.<\/jats:p>","DOI":"10.3233\/faia230437","type":"book-chapter","created":{"date-parts":[[2023,9,29]],"date-time":"2023-09-29T09:13:35Z","timestamp":1695978815000},"source":"Crossref","is-referenced-by-count":0,"title":["A Dynamic Selective Parameter Sharing Mechanism Embedded with Multi-Level Reasoning Abstractions"],"prefix":"10.3233","author":[{"given":"Yan","family":"Liu","sequence":"first","affiliation":[{"name":"Shenzhen Key Lab of Digital&Intelligent Technologies&Systems, College of Computer Science and Software Engineering, Shenzhen University, China"},{"name":"State Key Laboratory of Integrated Services Networks, Xidian University, China"}]},{"given":"Ying","family":"He","sequence":"additional","affiliation":[{"name":"Shenzhen Key Lab of Digital&Intelligent Technologies&Systems, College of Computer Science and Software Engineering, Shenzhen University, China"},{"name":"State Key Laboratory of Integrated Services Networks, Xidian University, China"}]},{"given":"Zhong","family":"Ming","sequence":"additional","affiliation":[{"name":"Shenzhen Key Lab of Digital&Intelligent Technologies&Systems, College of Computer Science and Software Engineering, Shenzhen University, China"},{"name":"Guangdong Laboratory of Artificial Intelligence and Digital Economy (SZ), Shenzhen, China"}]},{"given":"F. Richard","family":"Yu","sequence":"additional","affiliation":[{"name":"Shenzhen Key Lab of Digital&Intelligent Technologies&Systems, College of Computer Science and Software Engineering, Shenzhen University, China"},{"name":"Guangdong Laboratory of Artificial Intelligence and Digital Economy (SZ), Shenzhen, China"}]}],"member":"7437","container-title":["Frontiers in Artificial Intelligence and Applications","ECAI 2023"],"original-title":[],"link":[{"URL":"https:\/\/ebooks.iospress.nl\/pdf\/doi\/10.3233\/FAIA230437","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2023,9,29]],"date-time":"2023-09-29T09:13:38Z","timestamp":1695978818000},"score":1,"resource":{"primary":{"URL":"https:\/\/ebooks.iospress.nl\/doi\/10.3233\/FAIA230437"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,9,28]]},"ISBN":["9781643684369","9781643684376"],"references-count":0,"URL":"https:\/\/doi.org\/10.3233\/faia230437","relation":{},"ISSN":["0922-6389","1879-8314"],"issn-type":[{"value":"0922-6389","type":"print"},{"value":"1879-8314","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,9,28]]}}}