{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,2,5]],"date-time":"2026-02-05T23:48:02Z","timestamp":1770335282556,"version":"3.49.0"},"publisher-location":"New York, NY, USA","reference-count":31,"publisher":"ACM","license":[{"start":{"date-parts":[[2023,10,16]],"date-time":"2023-10-16T00:00:00Z","timestamp":1697414400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2023,10,23]]},"DOI":"10.1145\/3565287.3610255","type":"proceedings-article","created":{"date-parts":[[2023,9,28]],"date-time":"2023-09-28T19:59:44Z","timestamp":1695931184000},"page":"161-170","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":6,"title":["Distributional-Utility Actor-Critic for Network Slice Performance Guarantee"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0009-0001-8564-0407","authenticated-orcid":false,"given":"Jingdi","family":"Chen","sequence":"first","affiliation":[{"name":"George Washington University, Washington DC, District of Columbia, United States"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-3010-8090","authenticated-orcid":false,"given":"Tian","family":"Lan","sequence":"additional","affiliation":[{"name":"George Washington University, Washington DC, District of Columbia, United States of America"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-6272-1933","authenticated-orcid":false,"given":"Nakjung","family":"Choi","sequence":"additional","affiliation":[{"name":"Nokia Bell Labs, Murray Hill, New Jersey, United States of America"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2023,10,16]]},"reference":[{"key":"e_1_3_2_1_1_1","first-page":"00","article-title":"03","volume":"6","author":"Alliance O-RAN","year":"2022","unstructured":"O-RAN Alliance . 2022 . 03 . O-RAN Architecture-Description. Version 6 . 00 . O-RAN Alliance. 2022.03. O-RAN Architecture-Description. Version 6.00.","journal-title":"O-RAN Architecture-Description. Version"},{"key":"e_1_3_2_1_2_1","unstructured":"O-RAN Alliance. 2022.07. O-RAN Near-Real-time RAN Intelligent Controller E2 Service Model (E2SM) RC. Version 1.02.  O-RAN Alliance. 2022.07. O-RAN Near-Real-time RAN Intelligent Controller E2 Service Model (E2SM) RC. Version 1.02."},{"key":"e_1_3_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.48550\/ARXIV.1804.08617"},{"key":"e_1_3_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.48550\/ARXIV.1707.06887"},{"key":"e_1_3_2_1_5_1","volume-title":"Neuro-dynamic programming","author":"Bertsekas Dimitri","unstructured":"Dimitri Bertsekas and John N Tsitsiklis . 1996. Neuro-dynamic programming . Athena Scientific . Dimitri Bertsekas and John N Tsitsiklis. 1996. Neuro-dynamic programming. Athena Scientific."},{"key":"e_1_3_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.comnet.2020.107380"},{"key":"e_1_3_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/mcom.101.2001120"},{"key":"e_1_3_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1109\/tnnls.2021.3082568"},{"key":"e_1_3_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1109\/INFOCOM.2019.8737481"},{"key":"e_1_3_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/TVT.2019.2963462"},{"key":"e_1_3_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/MCOM.2017.1600951"},{"key":"e_1_3_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.48550\/ARXIV.1710.02298"},{"key":"e_1_3_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.48550\/ARXIV.1803.00933"},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.3390\/app9112361"},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2021.3072435"},{"key":"e_1_3_2_1_16_1","volume-title":"Kingma and Jimmy Ba","author":"Diederik","year":"2014","unstructured":"Diederik P. Kingma and Jimmy Ba . 2014 . Adam : A Method for Stochastic Optimization. http:\/\/arxiv.org\/abs\/1412.6980 cite arxiv:1412.6980Comment: Published as a conference paper at the 3rd International Conference for Learning Representations, San Diego , 2015. Diederik P. Kingma and Jimmy Ba. 2014. Adam: A Method for Stochastic Optimization. http:\/\/arxiv.org\/abs\/1412.6980 cite arxiv:1412.6980Comment: Published as a conference paper at the 3rd International Conference for Learning Representations, San Diego, 2015."},{"key":"e_1_3_2_1_17_1","volume-title":"An Axiomatic Theory of Fairness. CoRR abs\/0906.0557","author":"Lan Tian","year":"2009","unstructured":"Tian Lan , David T. H. Kao , Mung Chiang , and Ashutosh Sabharwal . 2009. An Axiomatic Theory of Fairness. CoRR abs\/0906.0557 ( 2009 ). arXiv:0906.0557 http:\/\/arxiv.org\/abs\/0906.0557 Tian Lan, David T. H. Kao, Mung Chiang, and Ashutosh Sabharwal. 2009. An Axiomatic Theory of Fairness. CoRR abs\/0906.0557 (2009). arXiv:0906.0557 http:\/\/arxiv.org\/abs\/0906.0557"},{"key":"e_1_3_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1145\/3323679.3326516"},{"key":"e_1_3_2_1_19_1","unstructured":"Yongsheng Mei Hanhan Zhou and Tian Lan. 2023. ReMIX: Regret Minimization for Monotonic Value Function Factorization in Multiagent Reinforcement Learning. arXiv:2302.05593 [cs.LG]  Yongsheng Mei Hanhan Zhou and Tian Lan. 2023. ReMIX: Regret Minimization for Monotonic Value Function Factorization in Multiagent Reinforcement Learning. arXiv:2302.05593 [cs.LG]"},{"key":"e_1_3_2_1_20_1","unstructured":"Yongsheng Mei Hanhan Zhou Tian Lan Guru Venkataramani and Peng Wei. 2023. MAC-PO: Multi-Agent Experience Replay via Collective Priority Optimization. arXiv:2302.10418 [cs.LG]  Yongsheng Mei Hanhan Zhou Tian Lan Guru Venkataramani and Peng Wei. 2023. MAC-PO: Multi-Agent Experience Replay via Collective Priority Optimization. arXiv:2302.10418 [cs.LG]"},{"key":"e_1_3_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.48550\/ARXIV.1602.01783"},{"key":"e_1_3_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.48550\/ARXIV.2102.08159"},{"key":"e_1_3_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.21314\/JOR.2000.038"},{"key":"e_1_3_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1145\/3281411.3281435"},{"key":"e_1_3_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.48550\/ARXIV.1511.05952"},{"key":"e_1_3_2_1_26_1","volume-title":"Robust optimization. Ph. D. Dissertation","author":"Sim Melvyn","unstructured":"Melvyn Sim . 2004. Robust optimization. Ph. D. Dissertation . Massachusetts Institute of Technology . Melvyn Sim. 2004. Robust optimization. Ph. D. Dissertation. Massachusetts Institute of Technology."},{"key":"e_1_3_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.48550\/ARXIV.2102.07936"},{"key":"e_1_3_2_1_28_1","volume-title":"Policy gradient methods for reinforcement learning with function approximation. Advances in neural information processing systems 12","author":"Sutton Richard S","year":"1999","unstructured":"Richard S Sutton , David McAllester , Satinder Singh , and Yishay Mansour . 1999. Policy gradient methods for reinforcement learning with function approximation. Advances in neural information processing systems 12 ( 1999 ). Richard S Sutton, David McAllester, Satinder Singh, and Yishay Mansour. 1999. Policy gradient methods for reinforcement learning with function approximation. Advances in neural information processing systems 12 (1999)."},{"key":"e_1_3_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.48550\/ARXIV.1911.03618"},{"key":"e_1_3_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.48550\/ARXIV.1911.02140"},{"key":"e_1_3_2_1_31_1","first-page":"15757","article-title":"PAC: Assisted Value Factorization with Counterfactual Predictions in Multi-Agent Reinforcement Learning","volume":"35","author":"Zhou Hanhan","year":"2022","unstructured":"Hanhan Zhou , Tian Lan , and Vaneet Aggarwal . 2022 . PAC: Assisted Value Factorization with Counterfactual Predictions in Multi-Agent Reinforcement Learning . Advances in Neural Information Processing Systems 35 (2022), 15757 -- 15769 . Hanhan Zhou, Tian Lan, and Vaneet Aggarwal. 2022. PAC: Assisted Value Factorization with Counterfactual Predictions in Multi-Agent Reinforcement Learning. Advances in Neural Information Processing Systems 35 (2022), 15757--15769.","journal-title":"Advances in Neural Information Processing Systems"}],"event":{"name":"MobiHoc '23: Twenty-fourth International Symposium on Theory, Algorithmic Foundations, and Protocol Design for Mobile Networks and Mobile Computing","location":"Washington DC USA","acronym":"MobiHoc '23","sponsor":["SIGMOBILE ACM Special Interest Group on Mobility of Systems, Users, Data and Computing"]},"container-title":["Proceedings of the Twenty-fourth International Symposium on Theory, Algorithmic Foundations, and Protocol Design for Mobile Networks and Mobile Computing"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3565287.3610255","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T16:37:43Z","timestamp":1750178263000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3565287.3610255"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,10,16]]},"references-count":31,"alternative-id":["10.1145\/3565287.3610255","10.1145\/3565287"],"URL":"https:\/\/doi.org\/10.1145\/3565287.3610255","relation":{},"subject":[],"published":{"date-parts":[[2023,10,16]]},"assertion":[{"value":"2023-10-16","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}