{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,27]],"date-time":"2026-07-27T00:14:36Z","timestamp":1785111276335,"version":"3.55.0"},"reference-count":15,"publisher":"American Institute of Aeronautics and Astronautics (AIAA)","issue":"7","funder":[{"DOI":"10.13039\/100006353","name":"Charles Stark Draper Laboratory","doi-asserted-by":"publisher","id":[{"id":"10.13039\/100006353","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["arc.aiaa.org"],"crossmark-restriction":true},"short-container-title":["Journal of Aerospace Information Systems"],"published-print":{"date-parts":[[2021,7]]},"abstract":"<jats:p> A policy for six-degree-of-freedom docking maneuvers with rotating targets is developed through reinforcement learning and implemented as a feedback control law. Potential clients for satellite servicing and orbital debris objects are often rotating around a constant axis within their respective Earth orbits. In the context of such missions, reinforcement learning provides an appealing framework for robust, autonomous maneuvers in uncertain environments with low on-board computational cost. This work uses proximal policy optimization to produce a docking policy for rotating or nonrotating targets that is valid over a portion of the six-degree-of-freedom state space while striving to minimize performance and control costs. Experiments using the simulated Apollo transposition and docking maneuver with an induced spin in the lunar module exhibit the policy\u2019s capabilities and provide a comparison with standard optimal control techniques. Furthermore, specific challenges and workarounds, as well as a discussion on the benefits and disadvantages of reinforcement learning for docking policies, are discussed to facilitate future research. As such, this work will serve as a foundation for further investigation of learning-based control laws for spacecraft proximity operations in uncertain environments. <\/jats:p>","DOI":"10.2514\/1.i010914","type":"journal-article","created":{"date-parts":[[2021,4,29]],"date-time":"2021-04-29T04:32:16Z","timestamp":1619670736000},"page":"417-428","update-policy":"https:\/\/doi.org\/10.2514\/aiaa_crossmarkpolicy","source":"Crossref","is-referenced-by-count":30,"title":["Autonomous Six-Degree-of-Freedom Spacecraft Docking with Rotating Targets via Reinforcement Learning"],"prefix":"10.2514","volume":"18","author":[{"ORCID":"https:\/\/orcid.org\/0000-0003-2896-3100","authenticated-orcid":false,"given":"Charles E.","family":"Oestreich","sequence":"first","affiliation":[{"name":"Massachusetts Institute of Technology, Cambridge, Massachusetts 02139"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Richard","family":"Linares","sequence":"additional","affiliation":[{"name":"Massachusetts Institute of Technology, Cambridge, Massachusetts 02139"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-0797-2470","authenticated-orcid":false,"given":"Ravi","family":"Gondhalekar","sequence":"additional","affiliation":[{"name":"The Charles Stark Draper Laboratory, Inc., Cambridge, Massachusetts 02139"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"1387","reference":[{"key":"r4","first-page":"1","volume-title":"Lecture Notes in Control and Information Sciences","volume":"460","author":"Starek J. A.","year":"2016"},{"key":"r5","doi-asserted-by":"publisher","DOI":"10.5139\/IJASS.2010.11.3.206"},{"key":"r6","doi-asserted-by":"publisher","DOI":"10.2514\/1.47645"},{"key":"r7","doi-asserted-by":"publisher","DOI":"10.1109\/TCST.2014.2379639"},{"key":"r8","doi-asserted-by":"publisher","DOI":"10.1109\/TAES.2016.140406"},{"key":"r9","doi-asserted-by":"publisher","DOI":"10.2514\/1.G003152"},{"key":"r14","doi-asserted-by":"publisher","DOI":"10.1016\/j.asr.2019.12.030"},{"key":"r15","doi-asserted-by":"publisher","DOI":"10.1016\/j.actaastro.2020.01.007"},{"key":"r18","first-page":"48","volume-title":"Wireless Communications Design Handbook","volume":"1","author":"Perez R.","year":"1998"},{"key":"r21","doi-asserted-by":"publisher","DOI":"10.1145\/2558904"},{"key":"r22","volume-title":"Reinforcement Learning: An Introduction","author":"Sutton R. S.","year":"2018","edition":"2"},{"key":"r26","doi-asserted-by":"publisher","DOI":"10.1214\/aoms\/1177729694"},{"key":"r27","doi-asserted-by":"publisher","DOI":"10.2514\/2.5048"},{"key":"r28","volume-title":"CSM\/LM Spacecraft Operation Data Book, Volume 3: Mass Properties","year":"1969"},{"key":"r29","volume-title":"Apollo Operations Handbook, Block II Spacecraft, Volume 1: Spacecraft Description","year":"1969"}],"container-title":["Journal of Aerospace Information Systems"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/arc.aiaa.org\/doi\/pdf\/10.2514\/1.I010914","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2023,7,14]],"date-time":"2023-07-14T17:19:48Z","timestamp":1689355188000},"score":1,"resource":{"primary":{"URL":"https:\/\/arc.aiaa.org\/doi\/10.2514\/1.I010914"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2021,7]]},"references-count":15,"journal-issue":{"issue":"7","published-print":{"date-parts":[[2021,7]]}},"alternative-id":["10.2514\/1.I010914"],"URL":"https:\/\/doi.org\/10.2514\/1.i010914","relation":{},"ISSN":["1940-3151","2327-3097"],"issn-type":[{"value":"1940-3151","type":"print"},{"value":"2327-3097","type":"electronic"}],"subject":[],"published":{"date-parts":[[2021,7]]},"assertion":[{"value":"2020-10-13","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2021-02-16","order":1,"name":"revised","label":"Revised","group":{"name":"publication_history","label":"Publication History"}},{"value":"2021-02-17","order":2,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2021-04-28","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}