{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,8,26]],"date-time":"2026-08-26T18:21:22Z","timestamp":1787768482614,"version":"build-2784847793"},"reference-count":99,"publisher":"American Association for the Advancement of Science (AAAS)","issue":"117","license":[{"start":{"date-parts":[[2027,8,26]],"date-time":"2027-08-26T00:00:00Z","timestamp":1819238400000},"content-version":"vor","delay-in-days":365,"URL":"https:\/\/www.science.org\/content\/page\/science-licenses-journal-article-reuse"}],"funder":[{"DOI":"10.13039\/100000179","name":"NSF Office of the Director","doi-asserted-by":"publisher","award":["CMMI-1944722"],"award-info":[{"award-number":["CMMI-1944722"]}],"id":[{"id":"10.13039\/100000179","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100000179","name":"NSF Office of the Director","doi-asserted-by":"publisher","award":["FRR-2153854"],"award-info":[{"award-number":["FRR-2153854"]}],"id":[{"id":"10.13039\/100000179","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100000185","name":"Defense Advanced Research Projects Agency","doi-asserted-by":"publisher","award":["HR00112490425"],"award-info":[{"award-number":["HR00112490425"]}],"id":[{"id":"10.13039\/100000185","id-type":"DOI","asserted-by":"publisher"}]},{"name":"Wu-Tsai Human Performance Alliance"},{"DOI":"10.13039\/100020670","name":"Stanford Institute for Human-Centered AI","doi-asserted-by":"crossref","id":[{"id":"10.13039\/100020670","id-type":"DOI","asserted-by":"crossref"}]},{"name":"The Robotics and AI Institute"},{"name":"BAIR Humanoid Intelligence Center"},{"name":"Design of Robustly Implementable Autonomous and Intelligent Machines"}],"content-domain":{"domain":["www.science.org"],"crossmark-restriction":true},"short-container-title":["Sci. Robot."],"published-print":{"date-parts":[[2026,8,26]]},"abstract":"<jats:p>The humanlike form of humanoid robots uniquely positions them to achieve the agility and versatility in motor skills that humans have. Learning from human demonstrations offers a scalable approach to acquiring these capabilities. However, prior works either produced unnatural motions or relied on motion-specific tuning to achieve satisfactory naturalness. Furthermore, these methods are often motion or goal specific, lacking the versatility to compose diverse skills, especially when solving unseen tasks. We present BeyondMimic, a framework that scales to diverse motions and carries the versatility to compose them seamlessly in tackling unseen downstream tasks. A compact motion tracking formulation enables mastery of a wide range of highly agile behaviors, including aerial cartwheels, spin kicks, flip kicks, and sprinting, with a single setup and shared hyperparameters, all while achieving humanlike performance. Moving beyond the mere imitation of existing motions, we propose a unified latent diffusion model that empowers versatile goal specification, seamless task switching, and dynamic composition of these agile behaviors. Leveraging classifier guidance, a diffusion-specific technique for test-time optimization toward unseen objectives, our model extended its capability to solve downstream tasks never encountered during training, including motion inpainting, joystick teleoperation, and obstacle avoidance, and transferred these skills zero-shot to real hardware. Together, these components enable scalable acquisition of humanlike motor skills from human motion and motion synthesis that generalizes and adapts beyond the training setup.<\/jats:p>","DOI":"10.1126\/scirobotics.adx8924","type":"journal-article","created":{"date-parts":[[2026,8,26]],"date-time":"2026-08-26T17:58:19Z","timestamp":1787767099000},"update-policy":"https:\/\/doi.org\/10.34133\/aaas_crossmark","source":"Crossref","is-referenced-by-count":0,"title":["BeyondMimic: From motion tracking to versatile humanoid control via guided diffusion"],"prefix":"10.1126","volume":"11","author":[{"ORCID":"https:\/\/orcid.org\/0009-0008-4104-1374","authenticated-orcid":true,"given":"Qiayuan","family":"Liao","sequence":"first","affiliation":[{"name":"University of California, Berkeley, Berkeley, CA 94720, USA."}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-9252-2428","authenticated-orcid":true,"given":"Takara E.","family":"Truong","sequence":"additional","affiliation":[{"name":"Stanford University, Stanford, CA 94305, USA."}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0005-0714-3711","authenticated-orcid":true,"given":"Xiaoyu","family":"Huang","sequence":"additional","affiliation":[{"name":"University of California, Berkeley, Berkeley, CA 94720, USA."}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-0377-770X","authenticated-orcid":true,"given":"Yuman","family":"Gao","sequence":"additional","affiliation":[{"name":"University of California, Berkeley, Berkeley, CA 94720, USA."}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-4376-2403","authenticated-orcid":true,"given":"Guy","family":"Tevet","sequence":"additional","affiliation":[{"name":"Stanford University, Stanford, CA 94305, USA."}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-5346-3637","authenticated-orcid":true,"given":"Koushil","family":"Sreenath","sequence":"additional","affiliation":[{"name":"University of California, Berkeley, Berkeley, CA 94720, USA."}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-5926-0905","authenticated-orcid":true,"given":"C. Karen","family":"Liu","sequence":"additional","affiliation":[{"name":"Stanford University, Stanford, CA 94305, USA."}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"221","reference":[{"key":"e_1_3_2_2_2","doi-asserted-by":"publisher","DOI":"10.1007\/s10514-015-9479-3"},{"key":"e_1_3_2_3_2","doi-asserted-by":"publisher","DOI":"10.1109\/TRO.2023.3324580"},{"key":"e_1_3_2_4_2","doi-asserted-by":"crossref","unstructured":"S. Kajita F. Kanehiro K. Kaneko K. Fujiwara K. Harada K. Yokoi H. Hirukawa \u201cBiped walking pattern generation by using preview control of zero-moment point\u201d in 2003 IEEE International Conference on Robotics and Automation (IEEE 2003) pp. 1620\u20131626.","DOI":"10.1109\/ROBOT.2003.1241826"},{"key":"e_1_3_2_5_2","doi-asserted-by":"crossref","unstructured":"J. Pratt J. Carff S. Drakunov A. Goswami \u201cCapture point: A step toward humanoid push recovery\u201d in 2006 6th IEEE-RAS International Conference on Humanoid Robots (IEEE 2006) pp. 200\u2013207.","DOI":"10.1109\/ICHR.2006.321385"},{"key":"e_1_3_2_6_2","doi-asserted-by":"crossref","unstructured":"R. Deits R. Tedrake \u201cFootstep planning on uneven terrain with mixed-integer convex optimization\u201d in 2014 IEEE-RAS International Conference on Humanoid Robots (IEEE 2014) pp. 279\u2013286.","DOI":"10.1109\/HUMANOIDS.2014.7041373"},{"key":"e_1_3_2_7_2","doi-asserted-by":"crossref","unstructured":"H. Dai A. Valenzuela R. Tedrake \u201cWhole-body motion planning with centroidal dynamics and full kinematics\u201d in 2014 IEEE-RAS International Conference on Humanoid Robots (IEEE 2014) pp. 295\u2013302.","DOI":"10.1109\/HUMANOIDS.2014.7041375"},{"key":"e_1_3_2_8_2","doi-asserted-by":"crossref","unstructured":"A. Hereid C. M. Hubicki E. A. Cousineau J. W. Hurst A. D. Ames \u201cHybrid zero dynamics based multiple shooting optimization with applications to robotic walking\u201d in 2015 IEEE International Conference on Robotics and Automation (ICRA) (IEEE 2015) pp. 5734\u20135740.","DOI":"10.1109\/ICRA.2015.7140002"},{"key":"e_1_3_2_9_2","doi-asserted-by":"crossref","unstructured":"J. Koenemann A. Del Prete Y. Tassa E. Todorov O. Stasse M. Bennewitz N. Mansard \u201cWhole-body model-predictive control applied to the HRP-2 humanoid\u201d in 2015 IEEE\/RSJ International Conference on Intelligent Robots and Systems (IROS) (IEEE 2015) pp. 3346\u20133351.","DOI":"10.1109\/IROS.2015.7353843"},{"key":"e_1_3_2_10_2","doi-asserted-by":"publisher","DOI":"10.1109\/JRA.1987.1087068"},{"key":"e_1_3_2_11_2","doi-asserted-by":"publisher","DOI":"10.1007\/s10514-015-9476-6"},{"key":"e_1_3_2_12_2","doi-asserted-by":"crossref","unstructured":"P. M. Wensing D. E. Orin \u201cGeneration of dynamic humanoid behaviors through task-space control with conic optimization\u201d in 2013 IEEE International Conference on Robotics and Automation (IEEE 2013) pp. 3103\u20133109.","DOI":"10.1109\/ICRA.2013.6631008"},{"key":"e_1_3_2_13_2","doi-asserted-by":"crossref","unstructured":"C. Khazoom D. Gonzalez-Diaz Y. Ding S. Kim \u201cHumanoid self-collision avoidance using whole-body control with control barrier functions\u201d in 2022 IEEE-RAS 21st International Conference on Humanoid Robots (Humanoids) (IEEE 2022) pp. 558\u2013565.","DOI":"10.1109\/Humanoids53995.2022.10000235"},{"key":"e_1_3_2_14_2","doi-asserted-by":"crossref","unstructured":"Z. Li B. Vanderborght N. G. Tsagarakis D. G. Caldwell \u201cQuasi-straightened knee walking for the humanoid robot\u201d in Modeling Simulation and Optimization of Bipedal Walking K. Mombaur K. Berns Eds. vol. 18 of Cognitive Systems Monograph (Springer 2013) pp. 117\u2013130.","DOI":"10.1007\/978-3-642-36368-9_9"},{"key":"e_1_3_2_15_2","doi-asserted-by":"crossref","unstructured":"J. Carpentier R. Budhiraja N. Mansard \u201cLearning feasibility constraints for multicontact locomotion of legged robots\u201d in Proceedings of Robotics: Science and Systems N. Amato S. Srinivasa N. Ayanian S. Kuindersma Eds. (RSS Foundation 2017) 10.15607\/RSS.2017.XIII.031.","DOI":"10.15607\/RSS.2017.XIII.031"},{"key":"e_1_3_2_16_2","doi-asserted-by":"crossref","unstructured":"S. Fasano J. Foster S. Bertrand C. DeBuys R. Griffin \u201cEfficient dynamic locomotion through step placement with straight legs and rolling contacts\u201d in 2024 IEEE International Conference on Robotics and Automation (ICRA) (IEEE 2024) pp. 1143\u20131150.","DOI":"10.1109\/ICRA57147.2024.10611056"},{"key":"e_1_3_2_17_2","doi-asserted-by":"publisher","DOI":"10.1142\/S0219843615500395"},{"key":"e_1_3_2_18_2","doi-asserted-by":"crossref","unstructured":"Y.-M. Chen G. Nelson R. Griffin M. Posa J. Pratt \u201cIntegrable whole-body orientation coordinates for legged robots\u201d in 2023 IEEE\/RSJ International Conference on Intelligent Robots and Systems (IROS) (IEEE 2023) pp. 10440\u201310447.","DOI":"10.1109\/IROS55552.2023.10341531"},{"key":"e_1_3_2_19_2","doi-asserted-by":"crossref","unstructured":"M. Chignoli D. Kim E. Stanger-Jones S. Kim \u201cThe MIT humanoid robot: Design motion planning and control for acrobatic behaviors\u201d in 2020 IEEE-RAS 20th International Conference on Humanoid Robots (Humanoids) (IEEE 2021) pp. 1\u20138.","DOI":"10.1109\/HUMANOIDS47582.2021.9555782"},{"key":"e_1_3_2_20_2","doi-asserted-by":"crossref","unstructured":"R. Subburaman N. G. Tsagarakis J. Lee \u201cOnline rolling motion generation for humanoid falls based on active energy control concepts\u201d in 2018 IEEE-RAS 18th International Conference on Humanoid Robots (Humanoids) (IEEE 2018) pp. 1\u20137.","DOI":"10.1109\/HUMANOIDS.2018.8624988"},{"key":"e_1_3_2_21_2","doi-asserted-by":"publisher","DOI":"10.1126\/scirobotics.adi9579"},{"key":"e_1_3_2_22_2","unstructured":"I. Radosavovic S. Kamat T. Darrell J. Malik Learning humanoid locomotion over challenging terrain. arXiv:2410.03654 [cs.RO] (2024)."},{"key":"e_1_3_2_23_2","doi-asserted-by":"crossref","unstructured":"Q. Liao B. Zhang X. Huang X. Huang Z. Li K. Sreenath \u201cBerkeley humanoid: A research platform for learning-based control\u201d in 2025 IEEE International Conference on Robotics and Automation (ICRA) (IEEE 2025) pp. 2897\u20132904.","DOI":"10.1109\/ICRA55743.2025.11127524"},{"key":"e_1_3_2_24_2","doi-asserted-by":"crossref","unstructured":"J. Long J. Ren M. Shi Z. Wang T. Huang P. Luo J. Pang \u201cLearning humanoid locomotion with perceptive internal model\u201d in 2025 IEEE International Conference on Robotics and Automation (ICRA) (IEEE 2025) pp. 9997\u201310003.","DOI":"10.1109\/ICRA55743.2025.11128333"},{"key":"e_1_3_2_25_2","doi-asserted-by":"crossref","unstructured":"J. Siekmann K. Green J. Warila A. Fern J. Hurst \u201cBlind bipedal stair traversal via sim-to-real reinforcement learning\u201d in Proceedings of Robotics: Science and Systems D. A. Shell M. Toussaint M. A. Hsieh Eds. (RSS Foundation 2021) 10.15607\/RSS.2021.XVII.061.","DOI":"10.15607\/RSS.2021.XVII.061"},{"key":"e_1_3_2_26_2","unstructured":"Z. Olkin K. Li W. D. Compton A. D. Ames Chasing stability: Humanoid running via control Lyapunov function guided reinforcement learning. arXiv:2509.19573 [cs.RO] (2025)."},{"key":"e_1_3_2_27_2","doi-asserted-by":"publisher","DOI":"10.1126\/scirobotics.adv3604"},{"key":"e_1_3_2_28_2","doi-asserted-by":"crossref","unstructured":"H. Wang Z. Wang J. Ren Q. Ben T. Huang W. Zhang J. Pang Beamdojo: Learning agile humanoid locomotion on sparse footholds. arXiv:2502.10363 [cs.RO] (2025).","DOI":"10.15607\/RSS.2025.XXI.068"},{"key":"e_1_3_2_29_2","unstructured":"P. Zhi P. Li J. Yin B. Jia S. Huang \u201cLearning a unified policy for position and force control in legged loco-manipulation\u201d in Proceedings of the 9th Conference on Robot Learning J. Lim S. Song H.-W. Park Eds. vol. 305 of Proceedings of Machine Learning Research (PMLR 2025) pp. 652\u2013669."},{"key":"e_1_3_2_30_2","doi-asserted-by":"publisher","DOI":"10.1109\/LRA.2025.3614078"},{"key":"e_1_3_2_31_2","unstructured":"P. Varin \u201cEstimation and planning for dynamic robot behaviors \u201d thesis Harvard University Cambridge MA (2021)."},{"key":"e_1_3_2_32_2","unstructured":"R. Deits S. Kuindersma M. P. Kelly T. Koolen Y. Abe B. Stephens US Patent 11 833 680 (2023)."},{"key":"e_1_3_2_33_2","doi-asserted-by":"publisher","DOI":"10.1145\/3197517.3201311"},{"key":"e_1_3_2_34_2","doi-asserted-by":"crossref","unstructured":"T. He J. Gao W. Xiao Y. Zhang Z. Wang J. Wang Z. Luo G. He N. Sobanbabu C. Pan Z. Yi G. Qu K. Kitani J. Hodgins L. J. Fan Y. Zhu C. Liu G. Shi Asap: Aligning simulation and real-world physics for learning agile humanoid whole-body skills. arXiv:2502.01143 [cs.RO] (2025).","DOI":"10.15607\/RSS.2025.XXI.066"},{"key":"e_1_3_2_35_2","unstructured":"T. Zhang B. Zheng R. Nai Y. Hu Y.-J. Wang G. Chen F. Lin J. Li C. Hong K. Sreenath Y. Gao HuB: Learning extreme humanoid balance. arXiv:2505.07294 [cs.RO] (2025)."},{"key":"e_1_3_2_36_2","doi-asserted-by":"crossref","unstructured":"W. Xie J. Han J. Zheng H. Li X. Liu J. Shi W. Zhang C. Bai X. Li KungfuBot: Physics-based humanoid whole-body control for learning highly-dynamic skills. arXiv:2506.12851 [cs.RO] (2025).","DOI":"10.52202\/085713-2089"},{"key":"e_1_3_2_37_2","doi-asserted-by":"publisher","DOI":"10.1145\/3450626.3459670"},{"key":"e_1_3_2_38_2","unstructured":"H. Wang W. Zhang R. Yu T. Huang J. Ren F. Jia Z. Wang X. Niu X. Chen J. Chen Q. Chen J. Wang J. Pang PhysHSI: Towards a real-world generalizable and natural humanoid-scene interaction system. arXiv:2510.11072 [cs.RO] (2025)."},{"key":"e_1_3_2_39_2","unstructured":"M. Ji X. Peng F. Liu J. Li G. Yang X. Cheng X. Wang Exbody2: Advanced expressive humanoid whole-body control. arXiv:2412.13196 [cs.RO] (2024)."},{"key":"e_1_3_2_40_2","unstructured":"Y. Ze Z. Chen J. P. Ara\u00fajo Z.-A. Cao X. B. Peng J. Wu C. K. Liu TWIST: Teleoperated whole-body imitation system. arXiv:2505.02833 [cs.RO] (2025)."},{"key":"e_1_3_2_41_2","unstructured":"Y. Li Y. Lin J. Cui T. Liu W. Liang Y. Zhu S. Huang CLONE: Closed-loop whole-body humanoid teleoperation for long-horizon tasks. arXiv:2506.08931 [cs.RO] (2025)."},{"key":"e_1_3_2_42_2","doi-asserted-by":"crossref","unstructured":"K. Yin W. Zeng K. Fan Z. Wang Q. Zhang Z. Tian J. Wang J. Pang W. Zhang UniTracker: Learning universal whole-body motion tracker for humanoid robots. arXiv:2507.07356 [cs.RO] (2025).","DOI":"10.1109\/LRA.2026.3692091"},{"key":"e_1_3_2_43_2","unstructured":"T. He Z. Luo X. He W. Xiao C. Zhang W. Zhang K. M. Kitani C. Liu G. Shi \u201cOmniH2O: Universal and dexterous human-to-humanoid whole-body teleoperation and learning\u201d in Proceedings of the 8th Conference on Robot Learning P. Agrawal O. Kroemer W. Burgard Eds. vol. 270 of Proceedings of Machine Learning Research (PMLR 2025) pp. 1516\u20131540."},{"key":"e_1_3_2_44_2","unstructured":"Z. Fu Q. Zhao Q. Wu G. Wetzstein C. Finn Humanplus: Humanoid shadowing and imitation from humans. arXiv:2406.10454 [cs.RO] (2024)."},{"key":"e_1_3_2_45_2","doi-asserted-by":"crossref","unstructured":"M. Xu Y. Shi K. Yin X. B. Peng \u201cParc: Physics-based augmentation with reinforcement learning for character controllers\u201d in SIGGRAPH Conference Papers \u201925: Proceedings of the Special Interest Group on Computer Graphics and Interactive Techniques Conference Conference Papers (Association for Computing Machinery 2025) pp. 1\u201311.","DOI":"10.1145\/3721238.3730616"},{"key":"e_1_3_2_46_2","doi-asserted-by":"crossref","unstructured":"X. Huang T. Truong Y. Zhang F. Yu J. P. Sleiman J. Hodgins K. Sreenath F. Farshidian Diffuse-CLoC: Guided diffusion for physics-based character look-ahead control. arXiv:2503.11801 [cs.GR] (2025).","DOI":"10.1145\/3731206"},{"key":"e_1_3_2_47_2","unstructured":"W. Zeng S. Lu K. Yin X. Niu M. Dai J. Wang J. Pang Behavior foundation model for humanoid robots. arXiv:2509.13780 [cs.RO] (2025)."},{"key":"e_1_3_2_48_2","doi-asserted-by":"crossref","unstructured":"Y. Shao X. Huang B. Zhang Q. Liao Y. Gao Y. Chi Z. Li S. Shao K. Sreenath LangWBC: Language-directed humanoid whole-body control via end-to-end learning. arXiv:2504.21738 [cs.RO] (2025).","DOI":"10.15607\/RSS.2025.XXI.065"},{"key":"e_1_3_2_49_2","doi-asserted-by":"crossref","unstructured":"Y. Wu K. Karunratanakul Z. Luo S. Tang UniPhys: Unified planner and controller with diffusion for flexible physics-based character control. arXiv:2504.12540 [cs.GR] (2025).","DOI":"10.1109\/ICCV51701.2025.01228"},{"key":"e_1_3_2_50_2","doi-asserted-by":"publisher","DOI":"10.1126\/scirobotics.aau5872"},{"key":"e_1_3_2_51_2","doi-asserted-by":"publisher","DOI":"10.1080\/14763141.2019.1574887"},{"key":"e_1_3_2_52_2","unstructured":"Z. Zhang J. Guo C. Chen J. Wang C. Lin Y. Lian H. Xue Z. Wang M. Liu J. Lyu H. Liu H. Wang L. Yi Track any motions under any disturbances. arXiv:2509.13833 [cs.RO] (2025)."},{"key":"e_1_3_2_53_2","unstructured":"Z. Chen M. Ji X. Cheng X. Peng X. B. Peng X. Wang GMT: General motion tracking for humanoid whole-body control. arXiv:2506.14770 [cs.RO] (2025)."},{"key":"e_1_3_2_54_2","unstructured":"K. Zakka B. Yi Q. Liao L. Le Lay MJLab: Isaac Lab API powered by MuJoCo-Warp for RL and robotics research GitHub (2025); https:\/\/github.com\/mujocolab\/mjlab."},{"key":"e_1_3_2_55_2","unstructured":"Unitree Robotics Unitree RL Lab: Reinforcement learning implementation for Unitree robots based on IsaacLab GitHub (2025); https:\/\/github.com\/unitreerobotics\/unitree_rl_lab."},{"key":"e_1_3_2_56_2","unstructured":"Unitree Robotics Unitree RL MJLab: Reinforcement learning for Unitree robots using MuJoCo GitHub (2025); https:\/\/github.com\/unitreerobotics\/unitree_rl_mjlab."},{"key":"e_1_3_2_57_2","unstructured":"Booster Robotics Booster Train: Reinforcement learning for Booster T1 humanoid robot GitHub (2025); https:\/\/github.com\/BoosterRobotics\/booster_train."},{"key":"e_1_3_2_58_2","unstructured":"Noetix Robotics Noetix E1 Lab: Reinforcement learning for Noetix E1 humanoid robot GitHub (2025); https:\/\/github.com\/Noetix-Robotics\/noetix_e1_lab."},{"key":"e_1_3_2_59_2","unstructured":"HighTorque Robotics Mini Pi Plus BeyondMimic: BeyondMimic implementation for HighTorque Mini Pi Plus humanoid robot GitHub (2025); https:\/\/github.com\/HighTorque-Robotics\/Mini-Pi-Plus_BeyondMimic."},{"key":"e_1_3_2_60_2","doi-asserted-by":"crossref","unstructured":"Z. Luo Y. Yuan T. Wang C. Li F. Casta\u00f1eda S. Chen Z.-A. Cao J. Li D. Minor Q. Ben J. Park D. Sami Z. Wang X. Da R. Ding C. Hogg L. Song E. Lim E. Jeong T. He H. Xue W. Xiao S. Yuen J. Kautz Y. Chang U. Iqbal L. Fan Y. Zhu SONIC: Supersizing motion tracking for natural humanoid whole-body control. arXiv:2511.07820 [cs.RO] (2025).","DOI":"10.1126\/scirobotics.aed4592"},{"key":"e_1_3_2_61_2","doi-asserted-by":"crossref","unstructured":"N. Ruiz Y. Li V. Jampani Y. Pritch M. Rubinstein K. Aberman \u201cDreambooth: Fine tuning text-to-image diffusion models for subject-driven generation\u201d in Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (IEEE 2023) pp. 22500\u201322510.","DOI":"10.1109\/CVPR52729.2023.02155"},{"key":"e_1_3_2_62_2","doi-asserted-by":"crossref","unstructured":"T. Brooks A. Holynski A. A. Efros \u201cInstructpix2pix: Learning to follow image editing instructions\u201d in Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (IEEE 2023) pp. 18392\u201318402.","DOI":"10.1109\/CVPR52729.2023.01764"},{"key":"e_1_3_2_63_2","unstructured":"A. Hertz R. Mokady J. Tenenbaum K. Aberman Y. Pritch D. Cohen-Or \u201cPrompt-to-prompt image editing with cross-attention control\u201d in The Eleventh International Conference on Learning Representations (OpenReview 2023); https:\/\/openreview.net\/forum?id=_CDixzkzeyb."},{"key":"e_1_3_2_64_2","doi-asserted-by":"crossref","unstructured":"L. Zhang A. Rao M. Agrawala \u201cAdding conditional control to text-to-image diffusion models\u201d in Proceedings of the IEEE\/CVF International Conference on Computer Vision (IEEE 2023) pp. 3836\u20133847.","DOI":"10.1109\/ICCV51070.2023.00355"},{"key":"e_1_3_2_65_2","doi-asserted-by":"crossref","unstructured":"C. Mou X. Wang L. Xie Y. Wu J. Zhang Z. Qi Y. Shan \u201cT2i-adapter: Learning adapters to dig out more controllable ability for text-to-image diffusion models\u201d in Proceedings of the AAAI Conference on Artificial Intelligence (AAAI 2024) vol. 38 pp. 4296\u20134304.","DOI":"10.1609\/aaai.v38i5.28226"},{"key":"e_1_3_2_66_2","unstructured":"Y. Wang S. Zhu J. Zhang J. Li Y. Li T. Liu S. Huang HDC: Humanoid Diffusion Controller GitHub (2025); https:\/\/humanoid-diffusion-controller.github.io\/ [accessed June 2025]."},{"key":"e_1_3_2_67_2","unstructured":"X. Huang Y. Chi R. Wang Z. Li X. B. Peng S. Shao B. Nikolic K. Sreenath \u201cDiffuseLoco: Real-time legged locomotion control with diffusion from offline datasets\u201d in Proceedings of the 8th Conference on Robot Learning P. Agrawal O. Kroemer W. Burgard Eds. vol. 270 of Proceedings of Machine Learning Research (PMLR 2024)."},{"key":"e_1_3_2_68_2","doi-asserted-by":"crossref","unstructured":"Y. Zhou C. Barnes J. Lu J. Yang H. Li \u201cOn the continuity of rotation representations in neural networks\u201d in Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (IEEE 2019) pp. 5745\u20135753.","DOI":"10.1109\/CVPR.2019.00589"},{"key":"e_1_3_2_69_2","doi-asserted-by":"crossref","unstructured":"Z. Luo J. Cao A. Winkler K. Kitani W. Xu \u201cPerpetual humanoid control for real-time simulated avatars\u201d in 2023 IEEE\/CVF International Conference on Computer Vision (ICCV) (IEEE 2023) pp. 10861\u201310870.","DOI":"10.1109\/ICCV51070.2023.01000"},{"key":"e_1_3_2_70_2","first-page":"6840","article-title":"Denoising diffusion probabilistic models","volume":"33","author":"Ho J.","year":"2020","unstructured":"J. Ho, A. Jain, P. Abbeel, Denoising diffusion probabilistic models. Adv. Neural Inf. Process. Syst. 33, 6840\u20136851 (2020).","journal-title":"Adv. Neural Inf. Process. Syst."},{"key":"e_1_3_2_71_2","unstructured":"A. Nichol P. Dhariwal A. Ramesh P. Shyam P. Mishkin B. McGrew I. Sutskever M. Chen Glide: Towards photorealistic image generation and editing with text-guided diffusion models. arXiv:2112.10741 [cs.CV] (2021)."},{"key":"e_1_3_2_72_2","unstructured":"M. Janner Y. Du J. B. Tenenbaum S. Levine Planning with diffusion for flexible behavior synthesis. arXiv:2205.09991 [cs.LG] (2022)."},{"key":"e_1_3_2_73_2","doi-asserted-by":"crossref","unstructured":"K. Karunratanakul K. Preechakul S. Suwajanakorn S. Tang \u201cGuided motion diffusion for controllable human motion synthesis\u201d in Proceedings of the IEEE\/CVF International Conference on Computer Vision (IEEE 2023) pp. 2151\u20132162.","DOI":"10.1109\/ICCV51070.2023.00205"},{"key":"e_1_3_2_74_2","doi-asserted-by":"crossref","unstructured":"S. Cohan G. Tevet D. Reda X. B. Peng M. van de Panne \u201cFlexible motion in-betweening with diffusion models\u201d in ACM SIGGRAPH 2024 Conference Papers (American Association for Computing Machinery 2024) pp. 1\u20139.","DOI":"10.1145\/3641519.3657414"},{"key":"e_1_3_2_75_2","doi-asserted-by":"crossref","unstructured":"H. Xue C. Pan Z. Yi G. Qu G. Shi \u201cFull-order sampling-based mpc for torque-level locomotion control via diffusion-style annealing\u201d in 2025 IEEE International Conference on Robotics and Automation (ICRA) (IEEE 2025) pp. 4974\u20134981.","DOI":"10.1109\/ICRA55743.2025.11127320"},{"key":"e_1_3_2_76_2","doi-asserted-by":"crossref","unstructured":"P. Roth J. Frey C. Cadena M. Hutter Learned perceptive forward dynamics model for safe and platform-aware robotic navigation. arXiv:2504.19322 [cs.RO] (2025).","DOI":"10.15607\/RSS.2025.XXI.001"},{"key":"e_1_3_2_77_2","unstructured":"M. Welling Y. W. Teh \u201cBayesian learning via stochastic gradient Langevin dynamics\u201d in Proceedings of the 28th International Conference on Machine Learning (ICML-11) (OmniPress 2011) pp. 681\u2013688."},{"key":"e_1_3_2_78_2","doi-asserted-by":"crossref","unstructured":"R. Rombach A. Blattmann D. Lorenz P. Esser B. Ommer \u201cHigh-resolution image synthesis with latent diffusion models\u201d in Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (IEEE 2022) pp. 10684\u201310695.","DOI":"10.1109\/CVPR52688.2022.01042"},{"key":"e_1_3_2_79_2","doi-asserted-by":"publisher","DOI":"10.1214\/aoms\/1177729694"},{"key":"e_1_3_2_80_2","unstructured":"Z. Luo J. Cao J. Merel A. Winkler J. Huang K. M. Kitani W. Xu \u201cUniversal humanoid motion representations for physics-based control\u201d in International Conference on Learning Representations B. Kim Y. Yue S. Chaudhuri K. Fragkiadaki M. Khan Y. Sun Eds. (ICLR 2024) pp. 56766\u201356782."},{"key":"e_1_3_2_81_2","unstructured":"S. Ross G. Gordon D. Bagnell \u201cA reduction of imitation learning and structured prediction to no-regret online learning\u201d in Proceedings of the Fourteenth International Conference on Artificial Intelligence and Statistics G. Gordon D. Dunson M. Dud\u00edk Eds. vol. 15 of Proceedings of Machine Learning Research (PMLR 2011) pp. 627\u2013635."},{"key":"e_1_3_2_82_2","unstructured":"Y. Song J. Sohl-Dickstein D. P. Kingma A. Kumar S. Ermon B. Poole \u201cScore-based generative modeling through stochastic differential equations\u201d in International Conference on Learning Representations (OpenReview 2021); https:\/\/openreview.net\/pdf\/ef0eadbe07115b0853e964f17aa09d811cd490f1.pdf."},{"key":"e_1_3_2_83_2","doi-asserted-by":"crossref","unstructured":"E. Todorov T. Erez Y. Tassa \u201cMuJoCo: A physics engine for model-based control\u201d in 2012 IEEE\/RSJ International Conference on Intelligent Robots and Systems (IEEE 2012) pp. 5026\u20135033.","DOI":"10.1109\/IROS.2012.6386109"},{"key":"e_1_3_2_84_2","doi-asserted-by":"publisher","DOI":"10.1111\/cgf.15175"},{"key":"e_1_3_2_85_2","doi-asserted-by":"crossref","unstructured":"V. B. Zordan J. K. Hodgins \u201cMotion capture-driven simulations that hit and react\u201d in Proceedings of the 2002 ACM SIGGRAPH\/Eurographics Symposium on Computer Animation (Association for Computing Machinery 2002) pp. 89\u201396.","DOI":"10.1145\/545261.545276"},{"key":"e_1_3_2_86_2","doi-asserted-by":"publisher","DOI":"10.1145\/1073204.1073249"},{"key":"e_1_3_2_87_2","doi-asserted-by":"crossref","unstructured":"J. P. Sleiman H. Li A. Adu-Bredu R. Deits A. Kumar K. Bergamin M. Bhardwaj S. Biddlestone N. Burger M. A. Estrada F. Iacobelli T. Koolen A. Lambert E. Lin M. E. Mungai Z. Nobles S. Rozen-Levy Y. Shi J. Wang J. Welner F. Yu M. Zhang A. Rizzi J. Hodgins S. Bertrand Y. Abe S. Kuindersma F. Farshidian ZEST: Zero-shot embodied skill transfer for athletic robot control. arXiv:2602.00401 [cs.RO] (2026).","DOI":"10.1126\/scirobotics.aec7695"},{"key":"e_1_3_2_88_2","doi-asserted-by":"publisher","DOI":"10.1145\/3386569.3392480"},{"key":"e_1_3_2_89_2","doi-asserted-by":"crossref","unstructured":"K. Fan S. Lu M. Dai R. Yu L. Xiao Z. Dou J. Dong L. Ma J. Wang Go to zero: Towards zero-shot motion generation with million-scale data. arXiv:2507.07095 [cs.CV] (2025).","DOI":"10.1109\/ICCV51701.2025.01239"},{"key":"e_1_3_2_90_2","doi-asserted-by":"crossref","unstructured":"T. E. Truong M. Piseno Z. Xie C. K. Liu \u201cPDP: Physics-based character animation via diffusion policy\u201d in SIGGRAPH Asia 2024 Conference Papers (ACM 2024) article no. 86.","DOI":"10.1145\/3680528.3687683"},{"key":"e_1_3_2_91_2","doi-asserted-by":"publisher","DOI":"10.1103\/PhysRev.36.823"},{"key":"e_1_3_2_92_2","doi-asserted-by":"crossref","unstructured":"R. Grandia F. Farshidian R. Ranftl M. Hutter Feedback MPC for torque-controlled legged robots. arXiv:1905.06144 [cs.RO] (2019).","DOI":"10.1109\/IROS40897.2019.8968251"},{"key":"e_1_3_2_93_2","unstructured":"J. Schulman F. Wolski P. Dhariwal A. Radford O. Klimov Proximal policy optimization algorithms. arXiv:1707.06347 [cs.LG] (2017)."},{"key":"e_1_3_2_94_2","unstructured":"M. Mittal P. Roth J. Tigue A. Richard O. Zhang P. Du A. Serrano-Mu\u00f1oz X. Yao R. Zurbr\u00fcgg N. Rudin L. Wawrzyniak M. Rakhsha A. Denzler E. Heiden A. Borovicka O. Ahmed I. Akinola A. Anwar M. T. Carlson J. Y. Feng A. Garg R. Gasoto L. Gulich Y. Guo M. Gussert A. Hansen M. Kulkarni C. Li W. Liu V. Makoviychuk G. Malczyk H. Mazhar M. Moghani A. Murali M. Noseworthy A. Poddubny N. Ratliff W. Rehberg C. Schwarke R. Singh J. L. Smith B. Tang R. Thaker M. Trepte K. Van Wyk F. Yu A. Millane V. Ramasamy R. Steiner S. Subramanian C. Volk C. Chen N. Jawale A. V. Kuruttukulam M. A. Lin A. Mandlekar K. Patzwaldt J. Welsh H. Zhao F. Anes J.-F. Lafleche N. Mo\u00ebnne-Loccoz S. Park R. Stepinski D. V. Gelder C. Amevor J. Carius J. Chang A. H. Chen P. de Heras Ciechomski G. Daviet M. Mohajerani J. von Muralt V. Reutskyy M. Sauter S. Schirm E. L. Shi P. Terdiman K. Vilella T. Widmer G. Yeoman T. Chen S. Grizan C. Li L. Li C. Smith R. Wiltz K. Alexis Y. Chang D. Chu L. J. Fan F. Farshidian A. Handa S. Huang M. Hutter Y. Narang S. Pouya S. Sheng Y. Zhu M. Macklin A. Moravanszky P. Reist Y. Guo D. Hoeller G. State Isaac Lab: A GPU-accelerated simulation framework for multi-modal robot learning. arXiv:2511.04831 [cs.RO] (2025)."},{"key":"e_1_3_2_95_2","doi-asserted-by":"crossref","unstructured":"T. Flayols A. Del Prete P. Wensing A. Mifsud M. Benallegue O. Stasse \u201cExperimental evaluation of simple estimators for humanoid robots\u201d in 2017 IEEE-RAS 17th International Conference on Humanoid Robotics (Humanoids) (IEEE 2017) pp. 889\u2013895.","DOI":"10.1109\/HUMANOIDS.2017.8246977"},{"key":"e_1_3_2_96_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.robot.2024.104750"},{"key":"e_1_3_2_97_2","unstructured":"Microsoft ONNX Runtime (2021); https:\/\/onnxruntime.ai\/."},{"key":"e_1_3_2_98_2","first-page":"5998","article-title":"Attention is all you need","volume":"30","author":"Vaswani A.","year":"2017","unstructured":"A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, \u0141. Kaiser, I. Polosukhin, Attention is all you need. Adv. Neural Inf. Process. Syst. 30, 5998\u20136008 (2017).","journal-title":"Adv. Neural Inf. Process. Syst."},{"key":"e_1_3_2_99_2","doi-asserted-by":"crossref","unstructured":"Z. Su X. Huang D. Ordo\u00f1ez-Apraez Y. Li Z. Li Q. Liao G. Turrisi M. Pontil C. Semini Y. Wu K. Sreenath \u201cLeveraging symmetry in rl-based legged locomotion control\u201d in 2024 IEEE\/RSJ International Conference on Intelligent Robots and Systems (IROS) (IEEE 2024) pp. 6899\u20136906.","DOI":"10.1109\/IROS58592.2024.10802439"},{"key":"e_1_3_2_100_2","unstructured":"B. M. Bell CppAD: A package for C++ algorithmic differentiation GitHub (2019); https:\/\/github.com\/coin-or\/CppAD\/tree\/stable\/20190200."}],"container-title":["Science Robotics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.science.org\/doi\/pdf\/10.1126\/scirobotics.adx8924","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/www.science.org\/doi\/pdf\/10.1126\/scirobotics.adx8924","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,8,26]],"date-time":"2026-08-26T17:58:36Z","timestamp":1787767116000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.science.org\/doi\/10.1126\/scirobotics.adx8924"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,8,26]]},"references-count":99,"journal-issue":{"issue":"117","published-print":{"date-parts":[[2026,8,26]]}},"alternative-id":["10.1126\/scirobotics.adx8924"],"URL":"https:\/\/doi.org\/10.1126\/scirobotics.adx8924","relation":{},"ISSN":["2470-9476"],"issn-type":[{"value":"2470-9476","type":"electronic"}],"subject":[],"published":{"date-parts":[[2026,8,26]]},"assertion":[{"value":"2025-11-06","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2026-07-31","order":2,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2026-08-26","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}],"article-number":"eadx8924"}}