{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,7,30]],"date-time":"2025-07-30T15:16:39Z","timestamp":1753888599129,"version":"3.41.2"},"reference-count":47,"publisher":"Wiley","issue":"1","license":[{"start":{"date-parts":[[2023,6,16]],"date-time":"2023-06-16T00:00:00Z","timestamp":1686873600000},"content-version":"vor","delay-in-days":166,"URL":"http:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"DOI":"10.13039\/501100002858","name":"China Postdoctoral Science Foundation","doi-asserted-by":"publisher","award":["2022M710921"],"award-info":[{"award-number":["2022M710921"]}],"id":[{"id":"10.13039\/501100002858","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["onlinelibrary.wiley.com"],"crossmark-restriction":true},"short-container-title":["International Journal of Intelligent Systems"],"published-print":{"date-parts":[[2023,1]]},"abstract":"<jats:p>Security issues are always considered in systems with wireless networks. However, few of them investigated covert signals existing on different communication channels to confuse advisories. In this paper, we consider the cooperation between two energy\u2010constrained agents, who could inject covert signals. First, the system performance is measured by Kullback\u2013Leibler divergence (KLD) to avoid much deviation. Then, the cooperative game between two agents is considered, in which two agents share the common goal at confusing advisories. More formally, this cooperative game is formulated as a Markov decision process (MDP) and the most economic strategies are obtained through reinforcement learning (RL) under the imperfect information. Finally, the feasibility of theoretical results is demonstrated on the interconnected New England test system (NETS) as well as its reduced system.<\/jats:p>","DOI":"10.1155\/2023\/9973580","type":"journal-article","created":{"date-parts":[[2023,6,17]],"date-time":"2023-06-17T01:05:06Z","timestamp":1686963906000},"update-policy":"https:\/\/doi.org\/10.1002\/crossmark_policy","source":"Crossref","is-referenced-by-count":0,"title":["Cooperative Game of Energy\u2010Constrained Agents in Wireless Communication Systems through Reinforcement Learning"],"prefix":"10.1155","volume":"2023","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-9822-6360","authenticated-orcid":false,"given":"Li","family":"Guo","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jianhong","family":"Wang","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Dian","family":"Huang","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Shengzhong","family":"Feng","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"311","published-online":{"date-parts":[[2023,6,16]]},"reference":[{"key":"e_1_2_10_1_2","doi-asserted-by":"publisher","DOI":"10.1109\/tc.2004.121"},{"key":"e_1_2_10_2_2","doi-asserted-by":"publisher","DOI":"10.1109\/tnet.2017.2690359"},{"key":"e_1_2_10_3_2","doi-asserted-by":"publisher","DOI":"10.1109\/mwc.01.1900525"},{"key":"e_1_2_10_4_2","doi-asserted-by":"publisher","DOI":"10.1109\/tsg.2014.2374577"},{"key":"e_1_2_10_5_2","doi-asserted-by":"crossref","unstructured":"CaoJ. ZhuX. andSunS. Age of loop oriented wireless networked control system: communication and control co-design in the FBL regime Proceedings of the IEEE INFOCOM 2022 - IEEE Conference on Computer Communications Workshops (INFOCOM WKSHPS) May 2022 New York NY USA.","DOI":"10.1109\/INFOCOMWKSHPS54753.2022.9798216"},{"key":"e_1_2_10_6_2","doi-asserted-by":"publisher","DOI":"10.1007\/s11432-018-9583-2"},{"key":"e_1_2_10_7_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.ins.2020.03.045"},{"key":"e_1_2_10_8_2","doi-asserted-by":"publisher","DOI":"10.1109\/jiot.2021.3073060"},{"key":"e_1_2_10_9_2","doi-asserted-by":"publisher","DOI":"10.1109\/tcns.2020.3035760"},{"key":"e_1_2_10_10_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.automatica.2017.09.028"},{"key":"e_1_2_10_11_2","doi-asserted-by":"publisher","DOI":"10.1109\/tcsi.2021.3071341"},{"key":"e_1_2_10_12_2","doi-asserted-by":"publisher","DOI":"10.1109\/tcns.2020.3024315"},{"key":"e_1_2_10_13_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.jnca.2012.12.017"},{"key":"e_1_2_10_14_2","doi-asserted-by":"publisher","DOI":"10.1002\/int.23063"},{"key":"e_1_2_10_15_2","doi-asserted-by":"publisher","DOI":"10.1002\/int.22397"},{"key":"e_1_2_10_16_2","doi-asserted-by":"publisher","DOI":"10.1109\/tsg.2018.2878570"},{"key":"e_1_2_10_17_2","first-page":"244","article-title":"Stealth false data injection using independent component analysis in smart grid","author":"Esmalifalak M.","year":"2011","journal-title":"Proc.IEEE Int. Conf. Smart Grid Commun."},{"key":"e_1_2_10_18_2","doi-asserted-by":"publisher","DOI":"10.1109\/tcns.2019.2910459"},{"key":"e_1_2_10_19_2","doi-asserted-by":"publisher","DOI":"10.1109\/tcst.2015.2462741"},{"key":"e_1_2_10_20_2","doi-asserted-by":"publisher","DOI":"10.1109\/tsg.2018.2813280"},{"key":"e_1_2_10_21_2","doi-asserted-by":"publisher","DOI":"10.1109\/tcst.2007.903066"},{"key":"e_1_2_10_22_2","doi-asserted-by":"crossref","unstructured":"WangJ. ZhangY. andKimT. Shapley q-value: a local reward approach to solve global reward games Proceedings of the 34th AAAI conf Artificial Intelligence February 2019 Hillton NY USA.","DOI":"10.1609\/aaai.v34i05.6220"},{"key":"e_1_2_10_23_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.neucom.2018.02.107"},{"key":"e_1_2_10_24_2","doi-asserted-by":"publisher","DOI":"10.1109\/tac.2017.2734840"},{"key":"e_1_2_10_25_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.automatica.2016.12.020"},{"volume-title":"Reinforcement Learning: An Introduction","year":"2017","author":"Sutton R.","key":"e_1_2_10_26_2"},{"key":"e_1_2_10_27_2","doi-asserted-by":"publisher","DOI":"10.1007\/bf00992696"},{"key":"e_1_2_10_28_2","doi-asserted-by":"publisher","DOI":"10.1109\/tcyb.2021.3108034"},{"key":"e_1_2_10_29_2","doi-asserted-by":"publisher","DOI":"10.1109\/tnnls.2022.3213566"},{"key":"e_1_2_10_30_2","doi-asserted-by":"publisher","DOI":"10.1109\/tnnls.2021.3098985"},{"key":"e_1_2_10_31_2","article-title":"Iterative solution of games by fictitious play","volume":"13","author":"Brown G. W.","year":"1951","journal-title":"Act. Anal. Prod Allocation."},{"key":"e_1_2_10_32_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.jet.2005.12.010"},{"key":"e_1_2_10_33_2","unstructured":"SilverD. LeverG. andHeessN. Deterministic policy gradient algorithms 32 Proceedings of the 31 st International Conference on Machine Learning June 2014 Beijing China."},{"key":"e_1_2_10_34_2","doi-asserted-by":"publisher","DOI":"10.1109\/tcns.2020.3030002"},{"key":"e_1_2_10_35_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.automatica.2022.110157"},{"key":"e_1_2_10_36_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.automatica.2022.110461"},{"key":"e_1_2_10_37_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.automatica.2022.110774"},{"key":"e_1_2_10_38_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-65048-3"},{"key":"e_1_2_10_39_2","doi-asserted-by":"crossref","unstructured":"LiuY. ZhangW. ChenF. andLiJ. Path planning based on improved deep deterministic policy gradient algorithm Proceedings of the 2019 IEEE 3rd Information Technology Networking Electronic and Automation Control Conference March 2019 Chengdu China.","DOI":"10.1109\/ITNEC.2019.8729369"},{"key":"e_1_2_10_40_2","doi-asserted-by":"publisher","DOI":"10.1007\/bf00992698"},{"volume-title":"Reinforcement Learning: An Introduction","year":"2018","author":"Sutton R.","key":"e_1_2_10_41_2"},{"key":"e_1_2_10_42_2","first-page":"1008","article-title":"Actor-critic algorithms","volume":"12","author":"Konda V.","year":"2000","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_2_10_43_2","first-page":"6379","article-title":"Multi-agent actor-critic for mixed cooperative-competitive environments","volume":"30","author":"Lowe R.","year":"2017","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_2_10_44_2","first-page":"1057","article-title":"Policy gradient methods for reinforcement learning with function approximation","volume":"12","author":"Sutton R.","year":"1999","journal-title":"Neural Info Processing Syst"},{"key":"e_1_2_10_45_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.geb.2005.08.005"},{"key":"e_1_2_10_46_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.jet.2020.105095"},{"volume-title":"Dynamic Estimation and Control of Power System","year":"2018","author":"Singh A.","key":"e_1_2_10_47_2"}],"container-title":["International Journal of Intelligent Systems"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/downloads.hindawi.com\/journals\/ijis\/2023\/9973580.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/downloads.hindawi.com\/journals\/ijis\/2023\/9973580.xml","content-type":"application\/xml","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/onlinelibrary.wiley.com\/doi\/pdf\/10.1155\/2023\/9973580","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,12,31]],"date-time":"2024-12-31T05:37:55Z","timestamp":1735623475000},"score":1,"resource":{"primary":{"URL":"https:\/\/onlinelibrary.wiley.com\/doi\/10.1155\/2023\/9973580"}},"subtitle":[],"editor":[{"given":"Vasudevan","family":"Rajamohan","sequence":"additional","affiliation":[],"role":[{"role":"editor","vocabulary":"crossref"}]}],"short-title":[],"issued":{"date-parts":[[2023,1]]},"references-count":47,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2023,1]]}},"alternative-id":["10.1155\/2023\/9973580"],"URL":"https:\/\/doi.org\/10.1155\/2023\/9973580","archive":["Portico"],"relation":{},"ISSN":["0884-8173","1098-111X"],"issn-type":[{"type":"print","value":"0884-8173"},{"type":"electronic","value":"1098-111X"}],"subject":[],"published":{"date-parts":[[2023,1]]},"assertion":[{"value":"2022-12-22","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2023-05-18","order":2,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2023-06-16","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}],"article-number":"9973580"}}