{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,3,10]],"date-time":"2026-03-10T21:38:55Z","timestamp":1773178735130,"version":"3.50.1"},"reference-count":26,"publisher":"Wiley","issue":"1","license":[{"start":{"date-parts":[[2021,6,2]],"date-time":"2021-06-02T00:00:00Z","timestamp":1622592000000},"content-version":"vor","delay-in-days":152,"URL":"http:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"content-domain":{"domain":["onlinelibrary.wiley.com"],"crossmark-restriction":true},"short-container-title":["Wireless Communications and Mobile Computing"],"published-print":{"date-parts":[[2021,1]]},"abstract":"<jats:p>In the adaptive traffic signal control (ATSC), reinforcement learning (RL) is a frontier research hotspot, combined with deep neural networks to further enhance its learning ability. The distributed multiagent RL (MARL) can avoid this kind of problem by observing some areas of each local RL in the complex plane traffic area. However, due to the limited communication capabilities between each agent, the environment becomes partially visible. This paper proposes multiagent reinforcement learning based on cooperative game (CG\u2010MARL) to design the intersection as an agent structure. The method considers not only the communication and coordination between agents but also the game between agents. Each agent observes its own area to learn the RL strategy and value function, then concentrates the <jats:italic>Q<\/jats:italic> function from different agents through a hybrid network, and finally forms its own final <jats:italic>Q<\/jats:italic> function in the entire large\u2010scale transportation network. The results show that the proposed method is superior to the traditional control method.<\/jats:p>","DOI":"10.1155\/2021\/6693636","type":"journal-article","created":{"date-parts":[[2021,6,2]],"date-time":"2021-06-02T22:36:24Z","timestamp":1622673384000},"update-policy":"https:\/\/doi.org\/10.1002\/crossmark_policy","source":"Crossref","is-referenced-by-count":5,"title":["Coordinated Control of Distributed Traffic Signal Based on Multiagent Cooperative Game"],"prefix":"10.1155","volume":"2021","author":[{"ORCID":"https:\/\/orcid.org\/0000-0001-7385-6377","authenticated-orcid":false,"given":"Zhenghua","family":"Zhang","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jin","family":"Qian","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Chongxin","family":"Fang","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Guoshu","family":"Liu","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Quan","family":"Su","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"311","published-online":{"date-parts":[[2021,6,2]]},"reference":[{"key":"e_1_2_8_1_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.tranpol.2017.07.002"},{"key":"e_1_2_8_2_2","doi-asserted-by":"publisher","DOI":"10.2307\/3006800"},{"key":"e_1_2_8_3_2","doi-asserted-by":"publisher","DOI":"10.1109\/ACC.1999.786344"},{"key":"e_1_2_8_4_2","doi-asserted-by":"publisher","DOI":"10.1049\/iet-its.2018.5308"},{"key":"e_1_2_8_5_2","doi-asserted-by":"publisher","DOI":"10.1109\/TEVC.2013.2260755"},{"key":"e_1_2_8_6_2","doi-asserted-by":"publisher","DOI":"10.1109\/TITS.2017.2762085"},{"key":"e_1_2_8_7_2","doi-asserted-by":"publisher","DOI":"10.1109\/TITS.2006.874716"},{"key":"e_1_2_8_8_2","doi-asserted-by":"publisher","DOI":"10.1007\/BF00992698"},{"key":"e_1_2_8_9_2","doi-asserted-by":"publisher","DOI":"10.1073\/pnas.88.18.8169"},{"key":"e_1_2_8_10_2","doi-asserted-by":"publisher","DOI":"10.3141\/1959-01"},{"key":"e_1_2_8_11_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.trpro.2015.09.070"},{"key":"e_1_2_8_12_2","doi-asserted-by":"publisher","DOI":"10.1049\/iet-its.2018.5170"},{"key":"e_1_2_8_13_2","doi-asserted-by":"crossref","unstructured":"Ha-liP.andKeD. An intersection signal control method based on deep reinforcement learning 2017 10th International Conference on Intelligent Computation Technology and Automation (ICICTA) 2017 Changsha China 344\u2013348 https:\/\/doi.org\/10.1109\/ICICTA.2017.83 2-s2.0-85047264229.","DOI":"10.1109\/ICICTA.2017.83"},{"key":"e_1_2_8_14_2","doi-asserted-by":"crossref","unstructured":"LiuY. LiuL. andChenW. Intelligent traffic light control using distributed multi-agent Q learning 2017 IEEE 20th International Conference on Intelligent Transportation Systems (ITSC) 2017 Yokohama 1\u20138 https:\/\/doi.org\/10.1109\/ITSC.2017.8317730 2-s2.0-85046266143.","DOI":"10.1109\/ITSC.2017.8317730"},{"key":"e_1_2_8_15_2","doi-asserted-by":"publisher","DOI":"10.1049\/iet-its.2009.0070"},{"key":"e_1_2_8_16_2","doi-asserted-by":"crossref","unstructured":"ChuT. WangJ. andCaoJ. Kernel-based reinforcement learning for traffic signal control with adaptive feature selection 53rd IEEE Conference on Decision and Control 2014 Los Angeles CA 1277\u20131282 https:\/\/doi.org\/10.1109\/CDC.2014.7039557 2-s2.0-84988222564.","DOI":"10.1109\/CDC.2014.7039557"},{"key":"e_1_2_8_17_2","doi-asserted-by":"publisher","DOI":"10.1038\/nature14236"},{"key":"e_1_2_8_18_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.trc.2017.09.020"},{"key":"e_1_2_8_19_2","doi-asserted-by":"publisher","DOI":"10.1109\/JAS.2016.7508798"},{"key":"e_1_2_8_20_2","doi-asserted-by":"publisher","DOI":"10.1109\/access.2019.2907618"},{"key":"e_1_2_8_21_2","doi-asserted-by":"crossref","unstructured":"LamouikI. YahyaouyA. andSabriM. A. Smart multi-agent traffic coordinator for autonomous vehicles at intersections 2017 International Conference on Advanced Technologies for Signal and Image Processing (ATSIP) 2017 Fez Morocco 1\u20136 https:\/\/doi.org\/10.1109\/ATSIP.2017.8075564 2-s2.0-85035344396.","DOI":"10.1109\/ATSIP.2017.8075564"},{"key":"e_1_2_8_22_2","doi-asserted-by":"publisher","DOI":"10.1109\/tvt.2020.2997896"},{"key":"e_1_2_8_23_2","doi-asserted-by":"crossref","unstructured":"MedinaJ. C.andBenekohalR. F. Traffic signal control using reinforcement learning and the max-plus algorithm as a coordinating strategy 2012 15th International IEEE Conference on Intelligent Transportation Systems 2012 Anchorage AK USA 596\u2013601 https:\/\/doi.org\/10.1109\/ITSC.2012.6338911 2-s2.0-84871219349.","DOI":"10.1109\/ITSC.2012.6338911"},{"key":"e_1_2_8_24_2","doi-asserted-by":"publisher","DOI":"10.1109\/CEC.2013.6557865"},{"key":"e_1_2_8_25_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.jebo.2015.11.001"},{"key":"e_1_2_8_26_2","first-page":"4295","article-title":"QMIX: monotonic value function factorisation for deep multi-agent reinforcement learning","volume":"80","author":"Rashid T.","year":"2018","journal-title":"International Conference of Machine Learning"}],"container-title":["Wireless Communications and Mobile Computing"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/downloads.hindawi.com\/journals\/wcmc\/2021\/6693636.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/downloads.hindawi.com\/journals\/wcmc\/2021\/6693636.xml","content-type":"application\/xml","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/onlinelibrary.wiley.com\/doi\/pdf\/10.1155\/2021\/6693636","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,8,7]],"date-time":"2024-08-07T12:23:20Z","timestamp":1723033400000},"score":1,"resource":{"primary":{"URL":"https:\/\/onlinelibrary.wiley.com\/doi\/10.1155\/2021\/6693636"}},"subtitle":[],"editor":[{"given":"Zhipeng","family":"Cai","sequence":"additional","affiliation":[],"role":[{"role":"editor","vocabulary":"crossref"}]}],"short-title":[],"issued":{"date-parts":[[2021,1]]},"references-count":26,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2021,1]]}},"alternative-id":["10.1155\/2021\/6693636"],"URL":"https:\/\/doi.org\/10.1155\/2021\/6693636","archive":["Portico"],"relation":{},"ISSN":["1530-8669","1530-8677"],"issn-type":[{"value":"1530-8669","type":"print"},{"value":"1530-8677","type":"electronic"}],"subject":[],"published":{"date-parts":[[2021,1]]},"assertion":[{"value":"2020-11-07","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2021-05-19","order":2,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2021-06-02","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}],"article-number":"6693636"}}