{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,25]],"date-time":"2026-07-25T10:55:43Z","timestamp":1784976943757,"version":"3.55.0"},"reference-count":213,"publisher":"Annual Reviews","issue":"1","content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Annu. Rev. Control Robot. Auton. Syst."],"published-print":{"date-parts":[[2020,5,3]]},"abstract":"<jats:p> In the context of robotics and automation, learning from demonstration (LfD) is the paradigm in which robots acquire new skills by learning to imitate an expert. The choice of LfD over other robot learning methods is compelling when ideal behavior can be neither easily scripted (as is done in traditional robot programming) nor easily defined as an optimization problem, but can be demonstrated. While there have been multiple surveys of this field in the past, there is a need for a new one given the considerable growth in the number of publications in recent years. This review aims to provide an overview of the collection of machine-learning methods used to enable a robot to learn from and imitate a teacher. We focus on recent advancements in the field and present an updated taxonomy and characterization of existing methods. We also discuss mature and emerging application areas for LfD and highlight the significant challenges that remain to be overcome both in theory and in practice. <\/jats:p>","DOI":"10.1146\/annurev-control-100819-063206","type":"journal-article","created":{"date-parts":[[2019,12,6]],"date-time":"2019-12-06T19:47:51Z","timestamp":1575661671000},"page":"297-330","source":"Crossref","is-referenced-by-count":697,"title":["Recent Advances in Robot Learning from Demonstration"],"prefix":"10.1146","volume":"3","author":[{"given":"Harish","family":"Ravichandar","sequence":"first","affiliation":[{"name":"Institute for Robotics and Intelligent Machines, Georgia Institute of Technology, Atlanta, Georgia 30332, USA;,"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Athanasios S.","family":"Polydoros","sequence":"additional","affiliation":[{"name":"Learning Algorithms and Systems Laboratory, \u00c9cole Polytechnique F\u00e9d\u00e9rale de Lausanne, 1015 Lausanne, Switzerland;,"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Sonia","family":"Chernova","sequence":"additional","affiliation":[{"name":"Institute for Robotics and Intelligent Machines, Georgia Institute of Technology, Atlanta, Georgia 30332, USA;,"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Aude","family":"Billard","sequence":"additional","affiliation":[{"name":"Learning Algorithms and Systems Laboratory, \u00c9cole Polytechnique F\u00e9d\u00e9rale de Lausanne, 1015 Lausanne, Switzerland;,"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"22","reference":[{"key":"B1","doi-asserted-by":"publisher","DOI":"10.1016\/S1364-6613(99)01327-3"},{"key":"B2","doi-asserted-by":"crossref","unstructured":"Billard A, Calinon S, Dillmann R, Schaal S. 2008. Robot programming by demonstration. In Springer Handbook of Robotics, ed. B Siciliano, O Khatib, pp. 1371\u201394. Berlin: Springer","DOI":"10.1007\/978-3-540-30301-5_60"},{"key":"B3","doi-asserted-by":"publisher","DOI":"10.1016\/j.robot.2008.10.024"},{"key":"B4","doi-asserted-by":"publisher","DOI":"10.2200\/S00568ED1V01Y201402AIM028"},{"key":"B5","doi-asserted-by":"crossref","unstructured":"Schulman J, Ho J, Lee A, Awwal I, Bradlow H, Abbeel P. 2013. Finding locally optimal, collision-free trajectories with sequential convex optimization. In Robotics: Science and Systems IX, ed. P Newman, D Fox, D Hsu, pap. 31. N.p.: Robot. Sci. Syst. Found.","DOI":"10.15607\/RSS.2013.IX.031"},{"key":"B6","doi-asserted-by":"publisher","DOI":"10.1177\/0278364913488805"},{"key":"B7","doi-asserted-by":"publisher","DOI":"10.3390\/robotics7020017"},{"key":"B8","doi-asserted-by":"publisher","DOI":"10.1109\/LRA.2017.2669369"},{"key":"B9","doi-asserted-by":"publisher","DOI":"10.1016\/S1367-5788(97)00014-X"},{"key":"B10","doi-asserted-by":"crossref","unstructured":"Billard AG, Calinon S, Dillmann R. 2016. Learning from humans. In Springer Handbook of Robotics, ed. B Siciliano, O Khatib, pp. 1995\u20132014. Berlin: Springer. 2nd ed.","DOI":"10.1007\/978-3-319-32552-1_74"},{"key":"B11","doi-asserted-by":"publisher","DOI":"10.1561\/2300000053"},{"key":"B12","doi-asserted-by":"publisher","DOI":"10.1109\/TRO.2013.2289018"},{"key":"B13","doi-asserted-by":"publisher","DOI":"10.1109\/HUMANOIDS.2016.7803328"},{"key":"B14","doi-asserted-by":"publisher","DOI":"10.1007\/s10514-016-9556-2"},{"key":"B15","doi-asserted-by":"publisher","DOI":"10.1007\/s11370-017-0235-8"},{"key":"B16","doi-asserted-by":"publisher","DOI":"10.1109\/LRA.2018.2833497"},{"key":"B17","doi-asserted-by":"publisher","DOI":"10.1109\/ROMAN.2017.8172424"},{"key":"B18","doi-asserted-by":"publisher","DOI":"10.1109\/HRI.2016.7451755"},{"key":"B19","doi-asserted-by":"publisher","DOI":"10.1109\/TSMCB.2006.886952"},{"key":"B20","doi-asserted-by":"crossref","unstructured":"Nehaniv CL, Dautenhahn K. 2002. The correspondence problem. In Imitation in Animals and Artifacts, ed. K Dautenhahn, CL Nehaniv, pp. 41\u201361. Cambridge, MA: MIT Press","DOI":"10.7551\/mitpress\/3676.001.0001"},{"key":"B21","doi-asserted-by":"publisher","DOI":"10.1177\/0278364910371999"},{"key":"B22","doi-asserted-by":"publisher","DOI":"10.1109\/ROBOT.2003.1242017"},{"key":"B23","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-28619-4_28"},{"key":"B24","doi-asserted-by":"publisher","DOI":"10.1145\/2696454.2696474"},{"key":"B25","doi-asserted-by":"crossref","unstructured":"Su Z, Kroemer O, Loeb GE, Sukhatme GS, Schaal S. 2016. Learning to switch between sensorimotor primitives using multimodal haptic signals. In International Conference on Simulation of Adaptive Behavior. pp. 170\u201382. Cham, Switz.: Springer","DOI":"10.1007\/978-3-319-43488-9_16"},{"key":"B26","doi-asserted-by":"publisher","DOI":"10.1163\/016918611X558261"},{"key":"B27","unstructured":"Rosen E, Whitney D, Phillips E, Ullman D, Tellex S. 2018. Testing robot teleoperation using a virtual reality interface with ROS reality. Paper presented at the 1st International Workshop on Virtual, Augmented, and Mixed Reality for Human-Robot Interaction, Chicago, Mar. 5"},{"key":"B28","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2018.8461249"},{"key":"B29","doi-asserted-by":"publisher","DOI":"10.1109\/IST.2018.8577081"},{"key":"B30","unstructured":"Whitney D, Rosen E, Tellex S. 2018. Learning from crowdsourced virtual reality demonstrations. Paper presented at the 1st International Workshop on Virtual, Augmented, and Mixed Reality for Human-Robot Interaction, Chicago, Mar. 5"},{"key":"B31","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2015.7139823"},{"key":"B32","unstructured":"Mandlekar A, Zhu Y, Garg A, Booher J, Spero M, et al. 2018. RoboTurk: a crowdsourcing platform for robotic skill learning through imitation. In Proceedings of the 2nd Conference on Robot Learning, ed. A Billared, A Dragan, J Peters, J Morimoto, pp. 879\u201393. Proc. Mach. Learn. Res. Vol. 87. N.p.: PMLR"},{"key":"B33","doi-asserted-by":"publisher","DOI":"10.1007\/s10514-015-9451-2"},{"key":"B34","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2011.5979632"},{"key":"B35","doi-asserted-by":"publisher","DOI":"10.1007\/s10514-018-9745-2"},{"key":"B36","doi-asserted-by":"publisher","DOI":"10.1109\/BIOROB.2018.8487959"},{"key":"B37","doi-asserted-by":"publisher","DOI":"10.1016\/j.robot.2004.03.005"},{"key":"B38","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2017.7989334"},{"key":"B39","doi-asserted-by":"publisher","DOI":"10.1109\/IROS.2014.6943191"},{"key":"B40","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2018.8460487"},{"key":"B41","doi-asserted-by":"publisher","DOI":"10.1109\/HUMANOIDS.2017.8246874"},{"key":"B42","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2018.8462901"},{"key":"B43","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-28872-7_20"},{"key":"B44","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-24586-7_9"},{"key":"B45","doi-asserted-by":"publisher","DOI":"10.1145\/2157689.2157693"},{"key":"B46","doi-asserted-by":"publisher","DOI":"10.1109\/TAMD.2010.2051030"},{"key":"B47","doi-asserted-by":"crossref","unstructured":"Bullard K, Schroecker Y, Chernova S. 2019. Active learning within constrained environments through imitation of an expert questioner. In Proceedings of the Twenty-Eighth International Joint Conference on Artificial Intelligence, ed. S Kraus, pp. 2045\u201352. Calif.: IJCAI","DOI":"10.24963\/ijcai.2019\/283"},{"key":"B48","doi-asserted-by":"publisher","DOI":"10.1109\/IROS.2018.8594279"},{"key":"B49","doi-asserted-by":"publisher","DOI":"10.1109\/HRI.2019.8673287"},{"key":"B50","doi-asserted-by":"publisher","DOI":"10.1145\/3171221.3171267"},{"key":"B51","doi-asserted-by":"publisher","DOI":"10.1609\/aimag.v35i4.2513"},{"key":"B52","doi-asserted-by":"publisher","DOI":"10.1109\/MIS.2017.3121552"},{"key":"B53","doi-asserted-by":"publisher","DOI":"10.1109\/HRI.2019.8673178"},{"key":"B54","first-page":"728","volume-title":"Proceedings of the 18th International Conference on Autonomous Agents and Multiagent Systems","author":"Kessler Faulkner T","year":"2019"},{"key":"B55","doi-asserted-by":"publisher","DOI":"10.1145\/3173386.3177066"},{"key":"B56","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2018.8461012"},{"key":"B57","unstructured":"Goodfellow I, Pouget-Abadie J, Mirza M, Xu B, Warde-Farley D, et al. 2014. Generative adversarial nets. In Advances in Neural Information Processing Systems 27, ed. Z Ghahramani, M Welling, C Cortes, ND Lawrence, KQ Weinberger, pp. 2672\u201380. Red Hook, NY: Curran"},{"key":"B58","unstructured":"Ho J, Ermon S. 2016. Generative adversarial imitation learning. In Advances in Neural Information Processing Systems 29, ed. DD Lee, M Sugiyama, UV Luxburg, I Guyon, R Garnett, pp. 4565\u201373. Red Hook, NY: Curran"},{"key":"B59","doi-asserted-by":"crossref","unstructured":"Schneider M, Ertel W. 2010. Robot learning by demonstration with local Gaussian process regression. 2010 IEEE\/RSJ International Conference on Intelligent Robots and Systems, pp. 255\u201360. Piscataway, NJ: IEEE","DOI":"10.1109\/IROS.2010.5650949"},{"key":"B60","doi-asserted-by":"publisher","DOI":"10.1007\/s12369-012-0160-0"},{"key":"B61","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2012.6225222"},{"key":"B62","unstructured":"Paraschos A, Daniel C, Peters J, Neumann G. 2013. Probabilistic movement primitives. In Advances in Neural Information Processing Systems 26, ed. CJC Burges, L Bottou, M Welling, Z Ghahramani, KQ Weinberger, pp. 2616\u201324. Red Hook, NY: Curran"},{"key":"B63","first-page":"1422","volume-title":"Proceedings of the Twenty-Seventh AAAI Conference on Artificial Intelligence","author":"Rozo L","year":"2013"},{"key":"B64","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2014.6907819"},{"key":"B65","doi-asserted-by":"publisher","DOI":"10.1109\/IROS.2014.6943190"},{"key":"B66","unstructured":"Calinon S, Evrard P. 2009. Learning collaborative manipulation tasks by demonstration using a haptic interface. In 2009 International Conference on Advanced Robotics. Piscataway, NJ: IEEE. https:\/\/ieeexplore.ieee.org\/document\/5174740"},{"key":"B67","doi-asserted-by":"publisher","DOI":"10.1109\/ROBOT.2009.5152385"},{"key":"B68","doi-asserted-by":"publisher","DOI":"10.1109\/IROS.2010.5648931"},{"key":"B69","doi-asserted-by":"publisher","DOI":"10.1109\/MRA.2010.936947"},{"key":"B70","doi-asserted-by":"publisher","DOI":"10.1109\/TRO.2011.2159412"},{"key":"B71","doi-asserted-by":"crossref","unstructured":"Rozo L, Jim\u00e9nez P, Torras C. 2011. Robot learning from demonstration of force-based tasks with multiple solution trajectories. IEEE 15th International Conference on Advanced Robotics, pp. 124\u201329. Piscataway, NJ: IEEE","DOI":"10.1109\/ICAR.2011.6088633"},{"key":"B72","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2012.6224877"},{"key":"B73","doi-asserted-by":"publisher","DOI":"10.1007\/s10514-013-9366-8"},{"key":"B74","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2014.6907861"},{"key":"B75","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2014.6907265"},{"key":"B76","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2015.7139639"},{"key":"B77","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2015.7138997"},{"key":"B78","doi-asserted-by":"publisher","DOI":"10.1016\/j.robot.2015.04.006"},{"key":"B79","doi-asserted-by":"publisher","DOI":"10.1109\/TMECH.2015.2510165"},{"key":"B80","doi-asserted-by":"publisher","DOI":"10.1016\/j.sysconle.2016.06.018"},{"key":"B81","unstructured":"Rana MA, Mukadam M, Ahmadzadeh SR, Chernova S, Boots B. 2017. Towards robust skill generalization: unifying learning from demonstration and motion planning. In Proceedings of the 1st Annual Conference on Robot Learning, ed. S Levine, V Vanhoucke, K Goldberg, pp. 109\u201318. Proc. Mach. Learn. Res. Vol. 78. N.p.: PMLR"},{"key":"B82","doi-asserted-by":"publisher","DOI":"10.1007\/s10514-018-9758-x"},{"key":"B83","doi-asserted-by":"publisher","DOI":"10.1109\/IROS.2018.8594103"},{"key":"B84","first-page":"527","volume-title":"2014 IEEE-RAS International Conference on Humanoid Robots","author":"Maeda G","year":"2015"},{"key":"B85","doi-asserted-by":"publisher","DOI":"10.1109\/ROBOT.2010.5509621"},{"key":"B86","doi-asserted-by":"publisher","DOI":"10.1109\/IROS.2007.4399227"},{"key":"B87","doi-asserted-by":"publisher","DOI":"10.1007\/s10514-015-9459-7"},{"key":"B88","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2015.7139553"},{"key":"B89","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2017.7989023"},{"key":"B90","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2017.7989324"},{"key":"B91","doi-asserted-by":"publisher","DOI":"10.1163\/156855308X360604"},{"key":"B92","doi-asserted-by":"publisher","DOI":"10.1109\/IROS.2017.8206344"},{"key":"B93","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2018.8461076"},{"key":"B94","unstructured":"Ravichandar H, Salehi I, Dani AP. 2017. Learning partially contracting dynamical systems from demonstrations. In Proceedings of the 1st Annual Conference on Robot Learning, ed. S Levine, V Vanhoucke, K Goldberg, pp. 369\u201378. Proc. Mach. Learn. Res. Vol. 78. N.p.: PMLR"},{"key":"B95","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2019.8793762"},{"key":"B96","doi-asserted-by":"publisher","DOI":"10.1109\/LRA.2018.2792531"},{"key":"B97","doi-asserted-by":"publisher","DOI":"10.1109\/IROS.2015.7353413"},{"key":"B98","doi-asserted-by":"publisher","DOI":"10.1016\/j.robot.2012.02.005"},{"key":"B99","doi-asserted-by":"publisher","DOI":"10.1162\/NECO_a_00393"},{"key":"B100","unstructured":"Umlauft J, Hirche S. 2017. Learning stable stochastic nonlinear dynamical systems. In Proceedings of the 34th International Conference on Machine Learning, ed. D Precup, YW Teh, pp. 3502\u201310. Proc. Mach. Learn. Res. Vol. 70. N.p.: PMLR"},{"key":"B101","doi-asserted-by":"publisher","DOI":"10.1109\/TRO.2018.2861921"},{"key":"B102","doi-asserted-by":"publisher","DOI":"10.1016\/j.ifacol.2019.01.001"},{"key":"B103","doi-asserted-by":"publisher","DOI":"10.1007\/s10514-017-9635-z"},{"key":"B104","doi-asserted-by":"publisher","DOI":"10.1109\/IROS.2016.7759153"},{"key":"B105","doi-asserted-by":"publisher","DOI":"10.1007\/s10846-017-0468-y"},{"key":"B106","doi-asserted-by":"publisher","DOI":"10.1177\/1059712313491614"},{"key":"B107","doi-asserted-by":"crossref","unstructured":"Ratliff N, Zucker M, Bagnell JA, Srinivasa S. 2009. CHOMP: gradient optimization techniques for efficient motion planning. In 2009 IEEE International Conference on Robotics and Automation, pp. 489\u201394. Piscataway, NJ: IEEE","DOI":"10.1109\/ROBOT.2009.5152817"},{"key":"B108","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2011.5980280"},{"key":"B109","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2013.6630743"},{"key":"B110","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2015.7139510"},{"key":"B111","unstructured":"Bajcsy A, Losey DP, O'Malley MK, Dragan AD. 2017. Learning robot objectives from physical human interaction. In Proceedings of the 1st Annual Conference on Robot Learning, ed. S Levine, V Vanhoucke, K Goldberg, pp. 217\u201326. Proc. Mach. Learn. Res. Vol. 78. N.p.: PMLR"},{"key":"B112","unstructured":"Ng AY, Russell SJ. 2000. Algorithms for inverse reinforcement learning. In Proceedings of the Seventeenth International Conference on Machine Learning, pp. 663\u201370. San Francisco: Morgan Kaufmann"},{"key":"B113","first-page":"335","volume-title":"Proceedings of the Twenty-Seventh International Conference on Machine Learning","author":"Dvijotham K","year":"2010"},{"key":"B114","unstructured":"Levine S, Koltun V. 2012. Continuous inverse optimal control with locally optimal examples. arXiv:1206.4617 [cs.LG]"},{"key":"B115","doi-asserted-by":"crossref","unstructured":"Doerr A, Ratliff ND, Bohg J, Toussaint M, Schaal S. 2015. Direct loss minimization inverse optimal control. In Robotics: Science and Systems XI, ed. LE Kavraki, D Hsu, J Buchli, pap. 13. N.p.: Robot. Sci. Syst. Found.","DOI":"10.15607\/RSS.2015.XI.013"},{"key":"B116","unstructured":"Finn C, Levine S, Abbeel P. 2016. Guided cost learning: deep inverse optimal control via policy optimization. In Proceedings of the 33rd International Conference on International Conference on Machine Learning, ed. MF Balcan, KQ Weinberger, pp. 49\u201358. Proc. Mach. Learn. Res. Vol. 48. N.p.: PMLR"},{"key":"B117","doi-asserted-by":"publisher","DOI":"10.1177\/0278364910369715"},{"key":"B118","unstructured":"Boularias A, Kober J, Peters J. 2011. Relative entropy inverse reinforcement learning. In Proceedings of the Fourteenth International Conference on Artificial Intelligence and Statistics, ed. G Gordon, D Dunson, M Dudik, pp. 182\u201389. Proc. Mach. Learn. Res. Vol. 15. N.p.: PMLR"},{"key":"B119","unstructured":"Hadfield-Menell D, Russell SJ, Abbeel P, Dragan A. 2016. Cooperative inverse reinforcement learning. In Advances in Neural Information Processing Systems 29, ed. DD Lee, M Sugiyama, UV Luxburg, I Guyon, R Garnett, pp. 3909\u201317. Red Hook, NY: Curran"},{"key":"B120","doi-asserted-by":"publisher","DOI":"10.1007\/s10514-009-9121-3"},{"key":"B121","doi-asserted-by":"crossref","unstructured":"Abbeel P, Ng AY. 2004. Apprenticeship learning via inverse reinforcement learning. In Proceedings of the Twenty-First International Conference on Machine Learning, pap. 1.New York: ACM","DOI":"10.1145\/1015330.1015430"},{"key":"B122","doi-asserted-by":"publisher","DOI":"10.1177\/0278364910392608"},{"key":"B123","unstructured":"Ziebart BD. 2010. Modeling purposeful adaptive behavior with the principle of maximum causal entropy. PhD Thesis, Carnegie Mellon Univ., Pittsburgh, PA"},{"key":"B124","unstructured":"Choi J, Kim KE. 2011. Map inference for Bayesian inverse reinforcement learning. In Advances in Neural Information Processing Systems 24, ed. J Shawe-Taylor, RS Zemel, PL Bartlett, F Pereira, KQ Weinberger, pp. 1989\u201397. Red Hook, NY: Curran"},{"key":"B125","doi-asserted-by":"publisher","DOI":"10.1109\/HRI.2019.8673256"},{"key":"B126","doi-asserted-by":"publisher","DOI":"10.1177\/0278364913495721"},{"key":"B127","doi-asserted-by":"publisher","DOI":"10.5772\/5611"},{"key":"B128","doi-asserted-by":"crossref","unstructured":"Grollman DH, Jenkins OC. 2010. Incremental learning of subtasks from unsegmented demonstration. 2010 IEEE\/RSJ International Conference on Intelligent Robots and Systems, pp. 261\u201366. Piscataway, NJ: IEEE","DOI":"10.1109\/IROS.2010.5650500"},{"key":"B129","doi-asserted-by":"publisher","DOI":"10.1007\/s10514-018-9749-y"},{"key":"B130","doi-asserted-by":"publisher","DOI":"10.1007\/s10514-018-9757-y"},{"key":"B131","doi-asserted-by":"publisher","DOI":"10.1177\/0278364911428653"},{"key":"B132","doi-asserted-by":"publisher","DOI":"10.1109\/IROS.2014.6943187"},{"key":"B133","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2016.7487760"},{"key":"B134","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2018.8460190"},{"key":"B135","doi-asserted-by":"publisher","DOI":"10.1109\/HUMANOIDS.2012.6651537"},{"key":"B136","doi-asserted-by":"publisher","DOI":"10.1109\/IROS.2010.5651244"},{"key":"B137","doi-asserted-by":"publisher","DOI":"10.1109\/IROS.2011.6094676"},{"key":"B138","doi-asserted-by":"publisher","DOI":"10.1177\/0278364911426178"},{"key":"B139","doi-asserted-by":"publisher","DOI":"10.1109\/IROS.2012.6386006"},{"key":"B140","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2014.6907441"},{"key":"B141","doi-asserted-by":"publisher","DOI":"10.1109\/HUMANOIDS.2015.7363584"},{"key":"B142","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2018.8461121"},{"key":"B143","doi-asserted-by":"publisher","DOI":"10.1109\/IROS.2015.7353415"},{"key":"B144","doi-asserted-by":"publisher","DOI":"10.1177\/0278364915587923"},{"key":"B145","doi-asserted-by":"publisher","DOI":"10.1145\/860575.860614"},{"key":"B146","doi-asserted-by":"publisher","DOI":"10.1109\/TSMCB.2006.886951"},{"key":"B147","doi-asserted-by":"publisher","DOI":"10.1007\/s12369-012-0162-y"},{"key":"B148","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2015.7139389"},{"key":"B149","doi-asserted-by":"publisher","DOI":"10.1613\/jair.1.11233"},{"key":"B150","doi-asserted-by":"publisher","DOI":"10.1109\/IROS.2010.5652268"},{"key":"B151","doi-asserted-by":"publisher","DOI":"10.1109\/ICHR.2010.5686284"},{"key":"B152","doi-asserted-by":"publisher","DOI":"10.1177\/0278364914554471"},{"key":"B153","unstructured":"Hausman K, Chebotar Y, Schaal S, Sukhatme G, Lim JJ. 2017. Multi-modal imitation learning from unstructured demonstrations using generative adversarial nets. In Advances in Neural Information Processing Systems 30, ed. I Guyon, UV Luxburg, S Bengio, H Wallach, R Fergus, et al., pp. 1235\u201345. Red Hook, NY: Curran"},{"key":"B154","doi-asserted-by":"publisher","DOI":"10.1177\/0278364918784350"},{"key":"B155","unstructured":"Krishnan S, Garg A, Liaw R, Miller L, Pokorny FT, Goldberg K. 2016. HIRL: hierarchical inverse reinforcement learning for long-horizon tasks with delayed rewards. arXiv:1604.06508 [cs.RO]"},{"key":"B156","unstructured":"Hirzinger G, Heindl J. 1983. Sensor programming, a new way for teaching a robot paths and forces torques simultaneously. In Proceedings of the 3rd International Conference on Robot Vision and Sensory Control, ed. B Rooks, pp. 549\u201358. Oxford, UK: Cotswold"},{"key":"B157","unstructured":"Asada H, Izumi H. 1987. Direct teaching and automatic program generation for the hybrid control of robot manipulators. In 1987 IEEE International Conference on Robotics and Automation, Vol. 4, 1401\u20136. Piscataway, NJ: IEEE"},{"key":"B158","doi-asserted-by":"publisher","DOI":"10.1016\/j.robot.2015.03.010"},{"key":"B159","doi-asserted-by":"publisher","DOI":"10.1109\/IROS.2014.6943028"},{"key":"B160","doi-asserted-by":"publisher","DOI":"10.1109\/LRA.2019.2894592"},{"key":"B161","doi-asserted-by":"crossref","unstructured":"Fong J, Tavakoli M. 2018. Kinesthetic teaching of a therapist's behavior to a rehabilitation robot. In 2018 International Symposium on Medical Robotics. Piscataway, NJ: IEEE. https:\/\/doi.org\/10.1109\/ISMR.2018.8333285","DOI":"10.1109\/ISMR.2018.8333285"},{"key":"B162","doi-asserted-by":"publisher","DOI":"10.1109\/LRA.2016.2521384"},{"key":"B163","doi-asserted-by":"publisher","DOI":"10.1007\/s41315-016-0006-2"},{"key":"B164","doi-asserted-by":"publisher","DOI":"10.1145\/3277903"},{"key":"B165","doi-asserted-by":"publisher","DOI":"10.1109\/TNSRE.2015.2501748"},{"key":"B166","doi-asserted-by":"publisher","DOI":"10.5898\/JHRI.2.1.Strabala"},{"key":"B167","doi-asserted-by":"publisher","DOI":"10.1162\/neco.1991.3.1.88"},{"key":"B168","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-33486-3_15"},{"key":"B169","unstructured":"Ross S, Gordon G, Bagnell D. 2011. A reduction of imitation learning and structured prediction to no-regret online learning. In Proceedings of the Fourteenth International Conference on Artificial Intelligence and Statistics, ed. G Gordon, D Dunson, M Dudik, pp. 627\u201335. Proc. Mach. Learn. Res. Vol. 15. N.p.: PMLR"},{"key":"B170","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2012.6224757"},{"key":"B171","unstructured":"Li Y, Song J, Ermon S. 2017. InfoGAIL: interpretable imitation learning from visual demonstrations. In Advances in Neural Information Processing Systems 30, ed. I Guyon, UV Luxburg, S Bengio, H Wallach, R Fergus, et al., pp. 3812\u201322. Red Hook, NY: Curran"},{"key":"B172","doi-asserted-by":"crossref","unstructured":"Pan Y, Cheng CA, Saigol K, Lee K, Yan X, et al. 2018. Agile autonomous driving using end-to-end deep imitation learning. In Robotics: Science and Systems XIV, ed. H Kress-Gazit, S Srinivasa, T Howard, N Atanasov, pap. 56. N.p.: Robot. Sci. Syst. Found.","DOI":"10.15607\/RSS.2018.XIV.056"},{"key":"B173","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2015.7139555"},{"key":"B174","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2013.6630809"},{"key":"B175","unstructured":"Kaufmann E, Loquercio A, Ranftl R, Dosovitskiy A, Koltun V, Scaramuzza D. 2018. Deep drone racing: learning agile flight in dynamic environments. In Proceedings of the 2nd Conference on Robot Learning, ed. A Billard, A Dragan, J Peters, J Morimoto, pp. 133\u201345. Proc. Mach. Learn. Res. Vol. 87. N.p.: PMLR"},{"key":"B176","doi-asserted-by":"publisher","DOI":"10.1109\/LRA.2018.2795643"},{"key":"B177","first-page":"39","volume-title":"Proceedings of the 2013 International Conference on Autonomous Agents and Multiagent Systems","author":"Farchy A","year":"2013"},{"key":"B178","first-page":"1594","volume-title":"Proceedings of the Twenty-Fourth AAAI Conference on Artificial Intelligence","author":"Meri\u00e7li\u00e7, Veloso M","year":"2010"},{"key":"B179","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-09584-4_25"},{"key":"B180","unstructured":"Kolter JZ, Abbeel P, Ng AY. 2008. Hierarchical apprenticeship learning with application to quadruped locomotion. In Advances in Neural Information Processing Systems 20, ed. JC Platt, D Koller, Y Singer, ST Roweis, pp. 769\u201376. Red Hook, NY: Curran"},{"key":"B181","doi-asserted-by":"publisher","DOI":"10.1177\/0278364910388677"},{"key":"B182","doi-asserted-by":"publisher","DOI":"10.1016\/j.robot.2004.03.003"},{"key":"B183","doi-asserted-by":"crossref","unstructured":"Carrera A, Palomeras N, Ribas D, Kormushev P, Carreras M. 2014. An intervention-AUV learns how to perform an underwater valve turning. In OCEANS 2014 - TAIPEI. Piscataway, NJ: IEEE. https:\/\/doi.org\/10.1109\/OCEANS-TAIPEI.2014.6964483","DOI":"10.1109\/OCEANS-TAIPEI.2014.6964483"},{"key":"B184","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2017.7989183"},{"key":"B185","doi-asserted-by":"publisher","DOI":"10.1109\/MRA.2018.2869523"},{"key":"B186","doi-asserted-by":"publisher","DOI":"10.1007\/s10514-015-9502-8"},{"key":"B187","unstructured":"Sun W, Venkatraman A, Gordon GJ, Boots B, Bagnell JA. 2017. Deeply AggreVaTeD: differentiable imitation learning for sequential prediction. In Proceedings of the 34th International Conference on Machine Learning, ed. D Precup, YW Teh, pp. 3309\u201318. Proc. Mach. Learn. Res. Vol. 70. N.p.: PMLR"},{"key":"B188","doi-asserted-by":"crossref","unstructured":"Kober J, Peters JR. 2009. Policy search for motor primitives in robotics. In Advances in Neural Information Processing Systems 21, ed. D Koller, D Schuurmans, Y Bengio, L Bottou, pp. 849\u201356. Red Hook, NY: Curran","DOI":"10.1109\/ROBOT.2009.5152577"},{"key":"B189","first-page":"617","volume-title":"The 10th International Conference on Autonomous Agents and Multiagent Systems","volume":"2","author":"Taylor ME","year":"2011"},{"key":"B190","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2011.5980200"},{"key":"B191","unstructured":"Kim B, Farahmand A, Pineau J, Precup D. 2013. Learning from limited demonstrations. In Advances in Neural Information Processing Systems 26, ed. CJC Burges, L Bottou, M Welling, Z Ghahramani, KQ Weinberger, pp. 2859\u201367. Red Hook, NY: Curran"},{"key":"B192","unstructured":"Vecerik M, Hester T, Scholz J, Wang F, Pietquin O, et al. 2017. Leveraging demonstrations for deep reinforcement learning on robotics problems with sparse rewards. arXiv:1707.08817 [cs.AI]"},{"key":"B193","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2019.8794074"},{"key":"B194","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2018.8463162"},{"key":"B195","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2018.8462891"},{"key":"B196","unstructured":"Brown DS, Niekum S. 2017. Toward probabilistic safety bounds for robot learning from demonstration. Tech. Rep. FS-17-01, Assoc. Adv. Artif. Intell., Palo Alto, CA"},{"key":"B197","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2016.7487167"},{"key":"B198","doi-asserted-by":"crossref","unstructured":"Zhou W, Li W. 2018. Safety-aware apprenticeship learning. In Computer Aided Verification: 30th International Conference, CAV 2018, ed. H Chockler, G Weissenbacher, pp. 662\u201380. Cham, Switz.: Springer","DOI":"10.1007\/978-3-319-96145-3_38"},{"key":"B199","doi-asserted-by":"publisher","DOI":"10.1109\/IROS.2016.7759557"},{"key":"B200","first-page":"5284","volume-title":"2013 IEEE International Conference on Robotics and Automation","author":"Ogrinc M","year":"2013"},{"key":"B201","doi-asserted-by":"publisher","DOI":"10.1109\/TASE.2016.2605707"},{"key":"B202","doi-asserted-by":"publisher","DOI":"10.1145\/1390156.1390175"},{"key":"B203","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2016.7487168"},{"key":"B204","doi-asserted-by":"publisher","DOI":"10.1126\/science.3629243"},{"key":"B205","doi-asserted-by":"publisher","DOI":"10.1006\/anbe.2003.2174"},{"key":"B206","unstructured":"Bagnell JA. 2015. An invitation to imitation. Tech. Rep. CMU-RI-TR-15-08, Robot. Inst., Carnegie Mellon Univ., Pittsburgh, PA"},{"key":"B207","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2014.6907339"},{"key":"B208","unstructured":"Finn C, Yu T, Zhang T, Abbeel P, Levine S. 2017. One-shot visual imitation learning via meta-learning. arXiv:1709.04905 [cs.LG]"},{"key":"B209","first-page":"27","volume-title":"Artificial Intelligence and Statistics 2001","author":"Corduneanu A","year":"2001"},{"key":"B210","doi-asserted-by":"publisher","DOI":"10.1002\/(SICI)1097-0266(199606)17:6<441::AID-SMJ819>3.0.CO;2-G"},{"key":"B211","doi-asserted-by":"publisher","DOI":"10.1016\/j.tics.2008.07.006"},{"key":"B212","doi-asserted-by":"crossref","unstructured":"Sung J, Jin SH, Saxena A. 2018. Robobarista: object part based transfer of manipulation trajectories from crowd-sourcing in 3D pointclouds. In Robotics Research, ed. A Bicchi, W Burgard, pp. 701\u201320. Cham, Switz.: Springer","DOI":"10.1007\/978-3-319-60916-4_40"},{"key":"B213","unstructured":"Castro PS, Li S, Zhang D. 2019. Inverse reinforcement learning with multiple ranked experts. arXiv:1907.13411 [cs.LG]"}],"container-title":["Annual Review of Control, Robotics, and Autonomous Systems"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.annualreviews.org\/doi\/pdf\/10.1146\/annurev-control-100819-063206","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2021,10,8]],"date-time":"2021-10-08T10:34:30Z","timestamp":1633689270000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.annualreviews.org\/doi\/10.1146\/annurev-control-100819-063206"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,5,3]]},"references-count":213,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2020,5,3]]}},"alternative-id":["10.1146\/annurev-control-100819-063206"],"URL":"https:\/\/doi.org\/10.1146\/annurev-control-100819-063206","relation":{},"ISSN":["2573-5144","2573-5144"],"issn-type":[{"value":"2573-5144","type":"print"},{"value":"2573-5144","type":"electronic"}],"subject":[],"published":{"date-parts":[[2020,5,3]]}}}