{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,4,29]],"date-time":"2025-04-29T04:45:08Z","timestamp":1745901908051,"version":"3.40.3"},"publisher-location":"Boston, MA","reference-count":31,"publisher":"Springer US","isbn-type":[{"type":"print","value":"9780387307688"},{"type":"electronic","value":"9780387301648"}],"license":[{"start":{"date-parts":[[2011,1,1]],"date-time":"2011-01-01T00:00:00Z","timestamp":1293840000000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/www.springernature.com\/gp\/researchers\/text-and-data-mining"},{"start":{"date-parts":[[2011,1,1]],"date-time":"2011-01-01T00:00:00Z","timestamp":1293840000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.springernature.com\/gp\/researchers\/text-and-data-mining"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2011]]},"DOI":"10.1007\/978-0-387-30164-8_45","type":"book-chapter","created":{"date-parts":[[2010,12,29]],"date-time":"2010-12-29T17:36:36Z","timestamp":1293644196000},"page":"53-61","source":"Crossref","is-referenced-by-count":3,"title":["Autonomous Helicopter Flight Using Reinforcement Learning"],"prefix":"10.1007","author":[{"given":"Adam","family":"Coates","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Pieter","family":"Abbeel","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Andrew Y.","family":"Ng","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","reference":[{"key":"45_CR1_45","doi-asserted-by":"crossref","unstructured":"Abbeel, P., Coates, A., Hunter, T., & Ng, A.\u00a0Y. (2008). Autonomous autorotation of an rc helicopter. In ISER 11.","DOI":"10.1007\/978-3-642-00196-3_45"},{"key":"45_CR2_45","doi-asserted-by":"crossref","unstructured":"Abbeel, P., Coates, A., Quigley, M., & Ng, A.\u00a0Y. (2007). An application of reinforcement learning to aerobatic helicopter flight. In NIPS 19 (pp. 1\u20138). Vancouver.","DOI":"10.7551\/mitpress\/7503.003.0006"},{"key":"45_CR3_45","unstructured":"Abbeel, P., Ganapathi, V., & Ng, A. Y. (2006). Learning vehicular dynamics with application to modeling helicopters. In NIPS 18. Vancouver."},{"key":"45_CR4_45","doi-asserted-by":"crossref","unstructured":"Abbeel, P., & Ng, A.\u00a0Y. (2004). Apprenticeship learning via inverse reinforcement learning. In Proceedings of the international conference on machine learning. New York: ACM.","DOI":"10.1145\/1015330.1015430"},{"key":"45_CR5_45","doi-asserted-by":"crossref","unstructured":"Abbeel, P., & Ng, A.\u00a0Y. (2005a). Exploration and apprenticeship learning in reinforcement learning. In Proceedings of the international conference on machine learning. New York: ACM","DOI":"10.1145\/1102351.1102352"},{"key":"45_CR6_45","unstructured":"Abbeel, P., & Ng, A.\u00a0Y. (2005b). Learning first order Markov models for control. In NIPS 18."},{"key":"45_CR7_45","first-page":"1","volume-title":"ICML \u201906: Proceedings of the 23rd international conference on machine learning","author":"P Abbeel","year":"2006","unstructured":"Abbeel, P., Quigley, M., & Ng, A.\u00a0Y. (2006). Using inaccurate models in reinforcement learning. In ICML \u201906: Proceedings of the 23rd international conference on machine learning (pp. 1\u20138). New York: ACM."},{"key":"45_CR8_45","volume-title":"Optimal control: linear quadratic methods","author":"B Anderson","year":"1989","unstructured":"Anderson, B., & Moore, J. (1989). Optimal control: linear quadratic methods. Princeton, NJ: Prentice-Hall."},{"key":"45_CR9_45","unstructured":"Bagnell, J., & Schneider, J. (2001). Autonomous helicopter control using reinforcement learning policy search methods. In International conference on robotics and automation. Canada: IEEE."},{"key":"45_CR10_45","first-page":"213","volume":"3","author":"RI Brafman","year":"2002","unstructured":"Brafman, R.\u00a0I., & Tennenholtz, M. (2002). R-max, a general polynomial time algorithm for near-optimal reinforcement learning. Journal of Machine Learning Research, 3, 213\u2013231.","journal-title":"Journal of Machine Learning Research"},{"key":"45_CR11_45","doi-asserted-by":"crossref","unstructured":"Coates, A., Abbeel, P., & Ng, A.\u00a0Y. (2008). Learning for control from multiple demonstrations. In ICML \u201908: Proceedings of the 25th international conference on machine learning.","DOI":"10.1145\/1390156.1390175"},{"key":"45_CR12_45","first-page":"1","volume":"391","author":"AP Dempster","year":"1977","unstructured":"Dempster, A.\u00a0P., Laird, N.\u00a0M., & Rubin, D.\u00a0B. (1977). Maximum likelihood from incomplete data via the EM algorithm. Journal of the Royal Statistical Society, 391, 1\u201338.","journal-title":"Journal of the Royal Statistical Society"},{"key":"45_CR13_45","doi-asserted-by":"crossref","unstructured":"Dunbabin, M., Brosnan, S., Roberts, J., & Corke, P. (2004). Vibration isolation for autonomous helicopter flight. In Proceedings of the IEEE international conference on robotics and automation (Vol.\u00a04, pp. 3609\u20133615).","DOI":"10.1109\/ROBOT.2004.1308812"},{"key":"45_CR14_45","doi-asserted-by":"crossref","unstructured":"Gavrilets, V., Martinos, I., Mettler, B., & Feron, E. (2002a). Control logic for automated aerobatic flight of miniature helicopter. In AIAA guidance, navigation and control conference. Cambridge, MA: Massachusetts Institute of Technology.","DOI":"10.2514\/6.2002-4834"},{"key":"45_CR15_45","unstructured":"Gavrilets, V., Martinos, I., Mettler, B., & Feron, E. (2002b). Flight test and simulation results for an autonomous aerobatic helicopter. In AIAA\/IEEE digital avionics systems conference."},{"key":"45_CR16_45","doi-asserted-by":"crossref","unstructured":"Gavrilets, V., Mettler, B., & Feron, E. (2001). Nonlinear model for a small-size acrobatic helicopter. In AIAA guidance, navigation and control conference (pp. 1593\u20131600).","DOI":"10.2514\/6.2001-4333"},{"key":"45_CR17_45","volume-title":"Differential dynamic programming","author":"DH Jacobson","year":"1970","unstructured":"Jacobson, D.\u00a0H., & Mayne, D. Q. (1970). Differential dynamic programming. New York: Elsevier."},{"key":"45_CR18_45","unstructured":"Kakade, S., Kearns, M., & Langford, J. (2003). Exploration in metric state spaces. In Proceedings of the international conference on machine learning."},{"key":"45_CR19_45","unstructured":"Kearns, M., & Koller, D. (1999). Efficient reinforcement learning in factored MDPs. In Proceedings of the 16th international joint conference on artificial intelligence. San Francisco: Morgan Kaufmann."},{"issue":"2\u20133","key":"45_CR20_45","doi-asserted-by":"crossref","first-page":"209","DOI":"10.1023\/A:1017984413808","volume":"49","author":"M Kearns","year":"2002","unstructured":"Kearns, M., & Singh, S. (2002). Near-optimal reinforcement learning in polynomial time. Machine Learning Journal, 49(2\u20133), 209\u2013232.","journal-title":"Machine Learning Journal"},{"key":"45_CR21_45","doi-asserted-by":"crossref","unstructured":"La Civita, M. (2003). Integrated modeling and robust control for full-envelope flight of robotic helicopters. PhD thesis, Carnegie Mellon University, Pittsburgh, PA.","DOI":"10.1109\/ROBOT.2003.1241652"},{"issue":"2","key":"45_CR22_45","doi-asserted-by":"crossref","first-page":"485","DOI":"10.2514\/1.15796","volume":"29","author":"M La Civita","year":"2006","unstructured":"La Civita, M., Papageorgiou, G., Messner, W.\u00a0C., & Kanade, T. (2006). Design and flight testing of a high-bandwidth $$\\mathcal{H}$$\n                \n                  \u221e\n                 loop shaping controller for a robotic helicopter. Journal of Guidance, Control, and Dynamics, 29(2), 485\u2013494.","journal-title":"Journal of Guidance, Control, and Dynamics"},{"key":"45_CR23_45","volume-title":"Principles of helicopter aerodynamics","author":"J Leishman","year":"2000","unstructured":"Leishman, J. (2000). Principles of helicopter aerodynamics. Cambridge: Cambridge University Press."},{"key":"45_CR24_45","doi-asserted-by":"crossref","first-page":"308","DOI":"10.1093\/comjnl\/7.4.308","volume":"7","author":"JA Nelder","year":"1964","unstructured":"Nelder, J. A., & Mead, R. (1964). A simplex method for function minimization. The Computer Journal, 7, 308\u2013313.","journal-title":"The Computer Journal"},{"key":"45_CR25_45","unstructured":"Ng, A.\u00a0Y., & Jordan, M. (2000). Pegasus: A policy search method for large MDPs and POMDPs. In Proceedings of the uncertainty in artificial intelligence 16th conference. San Francisco: Morgan Kaufmann."},{"key":"45_CR26_45","first-page":"663","volume-title":"Procedings of the 17th international conference on machine learning","author":"AY Ng","year":"2000","unstructured":"Ng, A. Y., & Russell, S. (2000). Algorithms for inverse reinforcement learning. In Procedings of the 17th international conference on machine learning (pp. 663\u2013670). San Francisco: Morgan Kaufmann."},{"key":"45_CR27_45","unstructured":"Ng, A. Y., Coates, A., Diel, M., Ganapathi, V., Schulte, J., Tse, B., et\u00a0al., (2004). Autonomous inverted helicopter flight via reinforcement learning. In International symposium on experimental robotics. Berlin: Springer."},{"key":"45_CR28_45","unstructured":"Ng, A. Y., Kim, H. J., Jordan, M., & Sastry, S. (2004). Autonomous helicopter flight via reinforcement learning. In NIPS 16."},{"issue":"3","key":"45_CR29_45","doi-asserted-by":"crossref","first-page":"371","DOI":"10.1109\/TRA.2003.810239","volume":"19","author":"S Saripalli","year":"2003","unstructured":"Saripalli, S., Montgomery, J. F., & Sukhatme, G. S. (2003). Visually-guided landing of an unmanned aerial vehicle. IEEE Transactions on Robotics and Autonomous Systems, 19(3), 371\u2013380.","journal-title":"IEEE Transactions on Robotics and Autonomous Systems"},{"key":"45_CR30_45","unstructured":"Seddon, J. (1990). Basic helicopter aerodynamics. In AIAA education series. El Segundo, CA: America Institute of Aeronautics and Astronautics."},{"key":"45_CR31_45","doi-asserted-by":"crossref","unstructured":"Tischler, M. B., & Cauffman, M. G. (1992). Frequency response method for rotorcraft system identification: Flight application to BO-105 couple rotor\/fuselage dynamics. Journal of the American Helicopter Society, 37.","DOI":"10.4050\/JAHS.37.3"}],"container-title":["Encyclopedia of Machine Learning"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/978-0-387-30164-8_45","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2023,12,22]],"date-time":"2023-12-22T03:03:59Z","timestamp":1703214239000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/978-0-387-30164-8_45"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2011]]},"ISBN":["9780387307688","9780387301648"],"references-count":31,"URL":"https:\/\/doi.org\/10.1007\/978-0-387-30164-8_45","relation":{},"subject":[],"published":{"date-parts":[[2011]]}}}