{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,3,13]],"date-time":"2026-03-13T08:58:19Z","timestamp":1773392299916,"version":"3.50.1"},"publisher-location":"New York, NY, USA","reference-count":13,"publisher":"ACM","license":[{"start":{"date-parts":[[2021,6,29]],"date-time":"2021-06-29T00:00:00Z","timestamp":1624924800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2021,6,29]]},"DOI":"10.1145\/3453892.3454004","type":"proceedings-article","created":{"date-parts":[[2021,6,29]],"date-time":"2021-06-29T17:01:49Z","timestamp":1624986109000},"page":"90-92","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":2,"title":["Accelerating Human-Agent Collaborative Reinforcement Learning"],"prefix":"10.1145","author":[{"given":"Fotios","family":"Lygerakis","sequence":"first","affiliation":[{"name":"University of Texas at Arlington National Center for Scientific Reseaarch Demokritos, United States"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Maria","family":"Dagioglou","sequence":"additional","affiliation":[{"name":"National Centre for Scientific Research Demokritos, Greece"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Vangelis","family":"Karkaletsis","sequence":"additional","affiliation":[{"name":"National Centre for Scientific Research Demokritos, Greece"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2021,6,29]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/DEVLRN.2008.4640845"},{"key":"e_1_3_2_1_2_1","unstructured":"Judith B\u00fctepage and Danica Kragic. 2017. Human-robot collaboration: from psychology to social robotics. arXiv preprint arXiv:1705.10146(2017).  Judith B\u00fctepage and Danica Kragic. 2017. Human-robot collaboration: from psychology to social robotics. arXiv preprint arXiv:1705.10146(2017)."},{"key":"#cr-split#-e_1_3_2_1_3_1.1","doi-asserted-by":"crossref","unstructured":"Giorgia Chiriatti Giacomo Palmieri and Matteo Palpacelli. 2020. A Framework for the Study of Human-Robot Collaboration in Rehabilitation Practices. 190-198. https:\/\/doi.org\/10.1007\/978-3-030-48989-2_21 10.1007\/978-3-030-48989-2_21","DOI":"10.1007\/978-3-030-48989-2_21"},{"key":"#cr-split#-e_1_3_2_1_3_1.2","doi-asserted-by":"crossref","unstructured":"Giorgia Chiriatti Giacomo Palmieri and Matteo Palpacelli. 2020. A Framework for the Study of Human-Robot Collaboration in Rehabilitation Practices. 190-198. https:\/\/doi.org\/10.1007\/978-3-030-48989-2_21","DOI":"10.1007\/978-3-030-48989-2_21"},{"key":"e_1_3_2_1_4_1","volume-title":"Advances in Neural Information Processing Systems, I.\u00a0Guyon, U.\u00a0V. Luxburg, S.\u00a0Bengio, H.\u00a0Wallach, R.\u00a0Fergus, S.\u00a0Vishwanathan, and R.\u00a0Garnett(Eds.), Vol.\u00a030. Curran Associates","author":"Christiano F","year":"2017","unstructured":"Paul\u00a0 F Christiano , Jan Leike , Tom Brown , Miljan Martic , Shane Legg , and Dario Amodei . 2017. Deep Reinforcement Learning from Human Preferences . In Advances in Neural Information Processing Systems, I.\u00a0Guyon, U.\u00a0V. Luxburg, S.\u00a0Bengio, H.\u00a0Wallach, R.\u00a0Fergus, S.\u00a0Vishwanathan, and R.\u00a0Garnett(Eds.), Vol.\u00a030. Curran Associates , Inc ., 4299\u20134307. https:\/\/proceedings.neurips.cc\/paper\/ 2017 \/file\/d5e2c0adad503c91f91df240d0cd4e49-Paper.pdf Paul\u00a0F Christiano, Jan Leike, Tom Brown, Miljan Martic, Shane Legg, and Dario Amodei. 2017. Deep Reinforcement Learning from Human Preferences. In Advances in Neural Information Processing Systems, I.\u00a0Guyon, U.\u00a0V. Luxburg, S.\u00a0Bengio, H.\u00a0Wallach, R.\u00a0Fergus, S.\u00a0Vishwanathan, and R.\u00a0Garnett(Eds.), Vol.\u00a030. Curran Associates, Inc., 4299\u20134307. https:\/\/proceedings.neurips.cc\/paper\/2017\/file\/d5e2c0adad503c91f91df240d0cd4e49-Paper.pdf"},{"key":"e_1_3_2_1_5_1","volume-title":"Soft Actor-Critic for Discrete Action Settings. (10","author":"Christodoulou Petros","year":"2019","unstructured":"Petros Christodoulou . 2019. Soft Actor-Critic for Discrete Action Settings. (10 2019 ). Petros Christodoulou. 2019. Soft Actor-Critic for Discrete Action Settings. (10 2019)."},{"key":"e_1_3_2_1_6_1","unstructured":"Tuomas Haarnoja Aurick Zhou Pieter Abbeel and Sergey Levine. 2018. Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor. CoRR abs\/1801.01290(2018). arxiv:1801.01290http:\/\/arxiv.org\/abs\/1801.01290  Tuomas Haarnoja Aurick Zhou Pieter Abbeel and Sergey Levine. 2018. Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor. CoRR abs\/1801.01290(2018). arxiv:1801.01290http:\/\/arxiv.org\/abs\/1801.01290"},{"key":"e_1_3_2_1_7_1","unstructured":"Tuomas Haarnoja Aurick Zhou Kristian Hartikainen George Tucker Sehoon Ha Jie Tan Vikash Kumar Henry Zhu Abhishek Gupta Pieter Abbeel and Sergey Levine. 2018. Soft Actor-Critic Algorithms and Applications. CoRR abs\/1812.05905(2018). arxiv:1812.05905http:\/\/arxiv.org\/abs\/1812.05905  Tuomas Haarnoja Aurick Zhou Kristian Hartikainen George Tucker Sehoon Ha Jie Tan Vikash Kumar Henry Zhu Abhishek Gupta Pieter Abbeel and Sergey Levine. 2018. Soft Actor-Critic Algorithms and Applications. CoRR abs\/1812.05905(2018). arxiv:1812.05905http:\/\/arxiv.org\/abs\/1812.05905"},{"key":"e_1_3_2_1_8_1","unstructured":"Wah\u00a0Loon Keng and Laura Graesser. 2017. SLM Lab. https:\/\/github.com\/kengz\/SLM-Lab.  Wah\u00a0Loon Keng and Laura Graesser. 2017. SLM Lab. https:\/\/github.com\/kengz\/SLM-Lab."},{"key":"e_1_3_2_1_9_1","volume-title":"Interactive","author":"Kragic Danica","unstructured":"Danica Kragic , Joakim Gustafson , Hakan Karaoguz , Patric Jensfelt , and Robert Krug . 2018. Interactive , Collaborative Robots : Challenges and Opportunities.. In IJCAI. 18\u201325. Danica Kragic, Joakim Gustafson, Hakan Karaoguz, Patric Jensfelt, and Robert Krug. 2018. Interactive, Collaborative Robots: Challenges and Opportunities.. In IJCAI. 18\u201325."},{"key":"e_1_3_2_1_10_1","volume-title":"Real-World Human-Robot Collaborative Reinforcement Learning. (03","author":"Shafti Ali","year":"2020","unstructured":"Ali Shafti , Jonas Tjomsland , William Dudley , and Aldo Faisal . 2020. Real-World Human-Robot Collaborative Reinforcement Learning. (03 2020 ). Ali Shafti, Jonas Tjomsland, William Dudley, and Aldo Faisal. 2020. Real-World Human-Robot Collaborative Reinforcement Learning. (03 2020)."},{"key":"e_1_3_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.mechatronics.2018.02.009"},{"key":"e_1_3_2_1_12_1","unstructured":"Garrett Warnell Nicholas\u00a0R. Waytowich Vernon Lawhern and Peter Stone. 2017. Deep TAMER: Interactive Agent Shaping in High-Dimensional State Spaces. CoRR abs\/1709.10163(2017). arxiv:1709.10163http:\/\/arxiv.org\/abs\/1709.10163  Garrett Warnell Nicholas\u00a0R. Waytowich Vernon Lawhern and Peter Stone. 2017. Deep TAMER: Interactive Agent Shaping in High-Dimensional State Spaces. CoRR abs\/1709.10163(2017). arxiv:1709.10163http:\/\/arxiv.org\/abs\/1709.10163"}],"event":{"name":"PETRA '21: The 14th PErvasive Technologies Related to Assistive Environments Conference","location":"Corfu Greece","acronym":"PETRA '21"},"container-title":["Proceedings of the 14th PErvasive Technologies Related to Assistive Environments Conference"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3453892.3454004","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3453892.3454004","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T20:17:41Z","timestamp":1750191461000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3453892.3454004"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2021,6,29]]},"references-count":13,"alternative-id":["10.1145\/3453892.3454004","10.1145\/3453892"],"URL":"https:\/\/doi.org\/10.1145\/3453892.3454004","relation":{},"subject":[],"published":{"date-parts":[[2021,6,29]]},"assertion":[{"value":"2021-06-29","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}