{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,2,21]],"date-time":"2025-02-21T15:25:28Z","timestamp":1740151528393,"version":"3.37.3"},"reference-count":14,"publisher":"Wiley","license":[{"start":{"date-parts":[[2013,4,18]],"date-time":"2013-04-18T00:00:00Z","timestamp":1366243200000},"content-version":"unspecified","delay-in-days":0,"URL":"http:\/\/creativecommons.org\/licenses\/by\/3.0\/"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Advances in Artificial Intelligence"],"published-print":{"date-parts":[[2013,4,18]]},"abstract":"<jats:p>We introduce a reinforcement learning architecture designed for problems with an infinite number of states, where each state can be seen as a vector of real numbers and with a finite number of actions, where each action requires a vector of real numbers as parameters. The main objective of this architecture is to distribute\nin two actors the work required to learn the final policy. One actor decides what action must be performed; meanwhile, a second actor determines the right parameters for the selected action. We tested our architecture and one algorithm based on it solving the robot dribbling problem, a challenging robot control problem taken from the RoboCup competitions. Our experimental work with three different function approximators provides enough evidence to prove that the proposed architecture can be used to implement fast, robust, and reliable reinforcement learning algorithms.<\/jats:p>","DOI":"10.1155\/2013\/492852","type":"journal-article","created":{"date-parts":[[2013,4,18]],"date-time":"2013-04-18T21:05:50Z","timestamp":1366319150000},"page":"1-10","source":"Crossref","is-referenced-by-count":1,"title":["A Novel Reinforcement Learning Architecture for Continuous State and Action Spaces"],"prefix":"10.1155","volume":"2013","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-4713-3762","authenticated-orcid":true,"given":"V\u00edctor","family":"Uc-Cetina","sequence":"first","affiliation":[{"name":"Facultad de Matem\u00e1ticas, Universidad Aut\u00f3noma de Yucat\u00e1n, Perif\u00e9rico Norte Tablaje 13615, Apartado Postal 192, C.P. 97119 M\u00e9rida, Yucat\u00e1n, Mexico"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"311","reference":[{"key":"3","doi-asserted-by":"publisher","DOI":"10.1016\/j.robot.2009.03.006"},{"key":"9","doi-asserted-by":"publisher","DOI":"10.1016\/j.asoc.2010.04.007"},{"volume-title":"Keepaway soccer: from machine learning testbed to benchmark","year":"2006","key":"16"},{"key":"10","doi-asserted-by":"publisher","DOI":"10.1016\/j.neucom.2010.11.012"},{"year":"1998","key":"18"},{"volume-title":"An actor\/critic algorithm that is equivalent to q-learning","year":"1995","key":"4"},{"key":"8","doi-asserted-by":"crossref","first-page":"237","DOI":"10.1613\/jair.301","volume":"4","year":"1996","journal-title":"Journal of Artificial Intelligence Research"},{"year":"1960","key":"7"},{"year":"1983","key":"14"},{"year":"1996","key":"2"},{"year":"1994","key":"11"},{"key":"1","doi-asserted-by":"publisher","DOI":"10.1023\/A:1022140919877"},{"key":"17","doi-asserted-by":"publisher","DOI":"10.1007\/BF00115009"},{"key":"15","doi-asserted-by":"publisher","DOI":"10.1023\/A:1007678930559"}],"container-title":["Advances in Artificial Intelligence"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/downloads.hindawi.com\/archive\/2013\/492852.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/downloads.hindawi.com\/archive\/2013\/492852.xml","content-type":"application\/xml","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/downloads.hindawi.com\/archive\/2013\/492852.pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2020,12,10]],"date-time":"2020-12-10T13:07:41Z","timestamp":1607605661000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.hindawi.com\/journals\/aai\/2013\/492852\/"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2013,4,18]]},"references-count":14,"alternative-id":["492852","492852"],"URL":"https:\/\/doi.org\/10.1155\/2013\/492852","relation":{},"ISSN":["1687-7470","1687-7489"],"issn-type":[{"type":"print","value":"1687-7470"},{"type":"electronic","value":"1687-7489"}],"subject":[],"published":{"date-parts":[[2013,4,18]]}}}