{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,26]],"date-time":"2026-07-26T09:43:24Z","timestamp":1785059004743,"version":"3.55.0"},"reference-count":69,"publisher":"SAGE Publications","issue":"10-11","license":[{"start":{"date-parts":[[2025,2,11]],"date-time":"2025-02-11T00:00:00Z","timestamp":1739232000000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/journals.sagepub.com\/page\/policies\/text-and-data-mining-license"}],"funder":[{"DOI":"10.13039\/501100010418","name":"Institute for Information and Communications Technology Promotion","doi-asserted-by":"publisher","award":["No.RS-2019-II19007, Artificial Intelligence Graduate School Program, KAIST"],"award-info":[{"award-number":["No.RS-2019-II19007, Artificial Intelligence Graduate School Program, KAIST"]}],"id":[{"id":"10.13039\/501100010418","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100003725","name":"National Research Foundation of Korea","doi-asserted-by":"publisher","award":["NRF-2021H1D3A2A03103683, Brain Pool Research Program"],"award-info":[{"award-number":["NRF-2021H1D3A2A03103683, Brain Pool Research Program"]}],"id":[{"id":"10.13039\/501100003725","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["journals.sagepub.com"],"crossmark-restriction":true},"short-container-title":["The International Journal of Robotics Research"],"published-print":{"date-parts":[[2025,9]]},"abstract":"<jats:p>Reinforcement learning (RL), imitation learning (IL), and task and motion planning (TAMP) have demonstrated impressive performance across various robotic manipulation tasks. However, these approaches have been limited to learning simple behaviors in current real-world manipulation benchmarks, such as pushing or pick-and-place. To enable more complex, long-horizon behaviors of an autonomous robot, we propose to focus on real-world furniture assembly, a complex, long-horizon robotic manipulation task that requires addressing many current robotic manipulation challenges. We present FurnitureBench, a reproducible real-world furniture assembly benchmark aimed at providing a low barrier for entry and being easily reproducible, so that researchers across the world can reliably test their algorithms and compare them against prior work. For ease of use, we provide 200+ hours of pre-collected data (5000+ demonstrations), 3D printable furniture models, a robotic environment setup guide, and systematic task initialization. Furthermore, we provide FurnitureSim, a fast and realistic simulator of FurnitureBench. We benchmark the performance of offline RL, IL, and offline-to-online RL algorithms on our assembly tasks and demonstrate the need to improve such algorithms to be able to solve our tasks in the real world, providing ample opportunities for future research.<\/jats:p>","DOI":"10.1177\/02783649241304789","type":"journal-article","created":{"date-parts":[[2025,2,11]],"date-time":"2025-02-11T12:08:27Z","timestamp":1739275707000},"page":"1863-1891","update-policy":"https:\/\/doi.org\/10.1177\/sage-journals-update-policy","source":"Crossref","is-referenced-by-count":10,"title":["FurnitureBench: Reproducible real-world benchmark for long-horizon complex manipulation"],"prefix":"10.1177","volume":"44","author":[{"ORCID":"https:\/\/orcid.org\/0009-0001-0351-2883","authenticated-orcid":false,"given":"Minho","family":"Heo","sequence":"first","affiliation":[{"name":"Korea Advanced Institute of Science and Technology"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-9918-1056","authenticated-orcid":false,"given":"Youngwoon","family":"Lee","sequence":"additional","affiliation":[{"name":"Yonsei University"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Doohyun","family":"Lee","sequence":"additional","affiliation":[{"name":"Korea Advanced Institute of Science and Technology"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0007-6891-676X","authenticated-orcid":false,"given":"Joseph J","family":"Lim","sequence":"additional","affiliation":[{"name":"Korea Advanced Institute of Science and Technology"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"179","published-online":{"date-parts":[[2025,2,11]]},"reference":[{"key":"e_1_3_5_2_1","unstructured":"Ahn M Zhu H Hartikainen K et al. (2020) Robel: robotics benchmarks for learning with low-cost robots. In: Conference on Robot Learning. Virtual November 16-18 2020."},{"key":"e_1_3_5_3_1","unstructured":"Akkaya I Andrychowicz M Chociej M et al. (2019) Solving rubik\u2019s cube with a robot hand. arXiv preprint arXiv:1910.07113."},{"key":"e_1_3_5_4_1","unstructured":"Ba JL Kiros JR Hinton GE (2016) Layer normalization. arXiv preprint arXiv:1607.06450."},{"key":"e_1_3_5_5_1","first-page":"1577","volume-title":"International Conference on Machine Learning","author":"Ball PJ","year":"2023","unstructured":"Ball PJ, Smith L, Kostrikov I, et al. (2023) Efficient online reinforcement learning with offline data. In: International Conference on Machine Learning. Westminster: PMLR, 1577\u20131594."},{"key":"e_1_3_5_6_1","first-page":"190","volume-title":"Neural Information Processing Systems 2021 Competitions and Demonstrations Track","author":"Bauer S","year":"2021","unstructured":"Bauer S, W\u00fcthrich M, Widmaier F, et al. (2021) Real robot challenge: a robotics competition in the cloud. In: Neural Information Processing Systems 2021 Competitions and Demonstrations Track. Westminster: PMLR, 190\u2013204."},{"key":"e_1_3_5_7_1","doi-asserted-by":"publisher","DOI":"10.1613\/jair.3912"},{"key":"e_1_3_5_8_1","doi-asserted-by":"publisher","DOI":"10.1109\/LRA.2020.2965865"},{"key":"e_1_3_5_9_1","unstructured":"Bradbury J Frostig R Hawkins P et al. (2018) JAX: composable transformations of Python+NumPy programs. https:\/\/github.com\/google\/jax."},{"key":"e_1_3_5_10_1","volume-title":"The OpenCV Library","author":"Bradski G","year":"2000","unstructured":"Bradski G (2000) The OpenCV Library. London: Dr. Dobb\u2019s Journal of Software Tools."},{"key":"e_1_3_5_11_1","unstructured":"Brockman G Cheung V Pettersson L et al. (2016) Openai gym. arXiv preprint arXiv:1606.01540."},{"key":"e_1_3_5_12_1","doi-asserted-by":"publisher","DOI":"10.15607\/RSS.2023.XIX.025"},{"key":"e_1_3_5_13_1","doi-asserted-by":"publisher","DOI":"10.1109\/LRA.2020.2972837"},{"key":"e_1_3_5_14_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2019.8793789"},{"key":"e_1_3_5_15_1","doi-asserted-by":"publisher","DOI":"10.15607\/RSS.2023.XIX.026"},{"key":"e_1_3_5_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/LRA.2020.2964160"},{"key":"e_1_3_5_17_1","first-page":"1087","volume-title":"Advances in Neural Information Processing Systems","author":"Duan Y","year":"2017","unstructured":"Duan Y, Andrychowicz M, Stadie B, et al. (2017) One-shot imitation learning. In: Advances in Neural Information Processing Systems. Cambridge: MIT Press, 1087\u20131098."},{"key":"e_1_3_5_18_1","doi-asserted-by":"crossref","unstructured":"Fu Z Zhao TZ Finn C (2024) Mobile aloha: learning bimanual mobile manipulation with low-cost whole-body teleoperation. In: Conference on Robot Learning Munich Germany November 6 to 9. 2024.","DOI":"10.15607\/RSS.2023.XIX.016"},{"key":"e_1_3_5_19_1","first-page":"20132","article-title":"A minimalist approach to offline reinforcement learning","volume":"34","author":"Fujimoto S","year":"2021","unstructured":"Fujimoto S, Gu SS (2021) A minimalist approach to offline reinforcement learning. Advances in Neural Information Processing Systems 34: 20132\u201320145.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_5_20_1","doi-asserted-by":"publisher","DOI":"10.1109\/LRA.2020.2965891"},{"key":"e_1_3_5_21_1","article-title":"Relay policy learning: solving long-horizon tasks via imitation and reinforcement learning","author":"Gupta A","year":"2019","unstructured":"Gupta A, Kumar V, Lynch C, et al. (2019) Relay policy learning: solving long-horizon tasks via imitation and reinforcement learning. Conference on Robot Learning.","journal-title":"Conference on Robot Learning"},{"key":"e_1_3_5_22_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-46493-0_38"},{"key":"e_1_3_5_23_1","doi-asserted-by":"publisher","DOI":"10.15607\/RSS.2023.XIX.041"},{"key":"e_1_3_5_24_1","first-page":"1","article-title":"Rlbench: the robot learning benchmark & learning environment","volume":"99","author":"James S","year":"2020","unstructured":"James S, Ma Z, Rovick Arrojo D, et al. (2020) Rlbench: the robot learning benchmark & learning environment. IEEE Robotics and Automation Letters 99: 1.","journal-title":"IEEE Robotics and Automation Letters"},{"key":"e_1_3_5_25_1","unstructured":"Jang E Irpan A Khansari M et al. (2021) Bc-z: zero-shot task generalization with robotic imitation learning. In: Conference on Robot Learning. London England November 8-11 2021."},{"key":"e_1_3_5_26_1","unstructured":"Jiang Y Wang C Zhang R et al. (2024) Transic: sim-to-real policy transfer by learning from online correction. In: Conference on Robot Learning Munich Germany Novmeber 6 to 9 2024."},{"key":"e_1_3_5_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/ROBOT.1992.220239"},{"key":"e_1_3_5_28_1","unstructured":"Kannan H Hafner D Finn C et al. (2021) Robodesk: a multi-task reinforcement learning benchmark. https:\/\/github.com\/google-research\/robodesk."},{"key":"e_1_3_5_29_1","doi-asserted-by":"publisher","DOI":"10.1109\/JRA.1987.1087068"},{"key":"e_1_3_5_30_1","doi-asserted-by":"publisher","DOI":"10.1109\/LRA.2020.2965869"},{"key":"e_1_3_5_31_1","unstructured":"Kingma DP Ba J (2015) Adam: a method for stochastic optimization. In: International Conference on Learning Representations San Diego May 7 to 9 2015."},{"key":"e_1_3_5_32_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2013.6630673"},{"key":"e_1_3_5_33_1","unstructured":"Kostrikov I Nair A Levine S (2022) Offline reinforcement learning with implicit q-learning. In: International Conference on Learning Representations Virtual April 25 to 29 2022. https:\/\/openreview.net\/forum?id=68n2s9ZJWF8."},{"key":"e_1_3_5_34_1","unstructured":"Lee Y Hu ES Yang Z et al (2019a) To follow or not to follow: selective imitation learning from observations. In: Conference on Robot Learning Osaka Japan October 29 to 31 2019."},{"key":"e_1_3_5_35_1","unstructured":"Lee AX Devin CM Zhou Y et al. (2021a) Beyond pick-and-place: tackling robotic stacking of diverse shapes. In: Conference on Robot Learning London England November 8 to 11 2021."},{"key":"e_1_3_5_36_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA48506.2021.9560986"},{"key":"e_1_3_5_37_1","unstructured":"Lee Y Sun SH Somasundaram S et al. (2019b) Composing complex skills by learning transition policies. In: International Conference on Learning Representations. New Orleans May 6 to 9 2019. Available at: https:\/\/openreview.net\/forum?id=rygrBhC5tQ."},{"issue":"1","key":"e_1_3_5_38_1","first-page":"1334","article-title":"End-to-end training of deep visuomotor policies","volume":"17","author":"Levine S","year":"2016","unstructured":"Levine S, Finn C, Darrell T, et al. (2016) End-to-end training of deep visuomotor policies. Journal of Machine Learning Research 17(1): 1334\u20131373.","journal-title":"Journal of Machine Learning Research"},{"key":"e_1_3_5_39_1","unstructured":"Li C Zhang R Wong J et al. (2022) Behavior-1k: a benchmark for embodied ai with 1 000 everyday activities and realistic simulation. In: Conference on Robot Learning Auckland New Zealand December 14 to 18 2022."},{"key":"e_1_3_5_40_1","unstructured":"Lin Y Wang AS Sutanto G et al. (2021) Polymetis. https:\/\/facebookresearch.github.io\/fairo\/polymetis\/."},{"key":"e_1_3_5_41_1","unstructured":"Loshchilov I Hutter F (2017) Decoupled weight decay regularization. arXiv preprint arXiv:1711.05101."},{"key":"e_1_3_5_42_1","unstructured":"Ma YJ Sodhani S Jayaraman D et al. (2022) Vip: towards universal visual reward and representation via value-implicit pre-training. arXiv preprint arXiv:2210.00030."},{"key":"e_1_3_5_43_1","unstructured":"Makoviychuk V Wawrzyniak L Guo Y et al. (2021) Isaac gym: high performance gpu based physics simulation for robot learning. In: Neural Information Processing Systems Datasets and Benchmarks Track Denver Colorado USA 10 Dec 2024 \u2013 Sun 15 Dec 2024."},{"key":"e_1_3_5_44_1","unstructured":"Mandlekar A Zhu Y Garg A et al. (2018) Roboturk: a crowdsourcing platform for robotic skill learning through imitation. In: Conference on Robot Learning Zurich Switzerland October 29 to 31 2018."},{"key":"e_1_3_5_45_1","unstructured":"Mandlekar A Xu D Wong J et al. (2021) What matters in learing from offline human demonstrations for robot manipulations. InLondon England November 8 to 11 2021."},{"key":"e_1_3_5_46_1","doi-asserted-by":"publisher","DOI":"10.1109\/LRA.2022.3180108"},{"key":"e_1_3_5_47_1","unstructured":"Nair S Rajeswaran A Kumar V et al. (2022) R3m: a universal visual representation for robot manipulation. In: Conference on Robot Learning Auckland New Zealand December 14 to 18 2022."},{"key":"e_1_3_5_48_1","volume-title":"Robotics: Science and Systems","author":"Narang Y","year":"2022","unstructured":"Narang Y, Storey K, Akinola I, et al. (2022) Factory: fast contact for robotic assembly. In: Robotics: Science and Systems. Cambridge: MIT Press."},{"key":"e_1_3_5_49_1","doi-asserted-by":"publisher","DOI":"10.15607\/RSS.2013.IX.048"},{"key":"e_1_3_5_50_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2011.5979561"},{"key":"e_1_3_5_51_1","first-page":"6892","volume-title":"Open x-embodiment: robotic learning datasets and rt-x models","author":"Open X-Embodiment Collaboration","year":"2024","unstructured":"Open X-Embodiment Collaboration, Padalkar A, Pooley A, Jain A, et al (2024) Open x-embodiment: robotic learning datasets and rt-x models. 2024 IEEE International Conference on Robotics and Automation. IEEE, 6892-6903."},{"key":"e_1_3_5_52_1","doi-asserted-by":"publisher","DOI":"10.1177\/0278364919887447"},{"key":"e_1_3_5_53_1","unstructured":"Park S Frans K Levine S et al. (2024) Is value learning really the main bottleneck in offline rl? arXiv preprint arXiv:2406: 09329."},{"key":"e_1_3_5_54_1","first-page":"305","volume-title":"Advances in Neural Information Processing Systems","author":"Pomerleau DA","year":"1989","unstructured":"Pomerleau DA (1989) Alvinn: an autonomous land vehicle in a neural network. In: Advances in Neural Information Processing Systems. Cambridge: MIT Press, 305\u2013313."},{"key":"e_1_3_5_55_1","doi-asserted-by":"publisher","DOI":"10.15607\/RSS.2018.XIV.049"},{"key":"e_1_3_5_56_1","doi-asserted-by":"publisher","DOI":"10.1108\/RPJ-03-2022-0081"},{"key":"e_1_3_5_57_1","unstructured":"Srivastava S Li C Lingelbach M et al. (2021) Behavior: benchmark for everyday household activities in virtual interactive and ecological environments. In Conference on Robot Learning London England November 8 to 11 2021."},{"key":"e_1_3_5_58_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2016.7487162"},{"key":"e_1_3_5_59_1","doi-asserted-by":"publisher","DOI":"10.1126\/scirobotics.aat6385"},{"key":"e_1_3_5_60_1","unstructured":"Szot A Clegg A Undersander E et al. (2021) Habitat 2.0: training home assistants to rearrange their habitat. Advances in Neural Information Processing Systems Virtual December 6 to 14 2021. Curran Associates Inc."},{"key":"e_1_3_5_61_1","unstructured":"Tassa Y Doron Y Muldal A et al. (2018) Deepmind control suite. arXiv preprint arXiv:1801.00690."},{"key":"e_1_3_5_62_1","unstructured":"Urakami Y Hodgkinson A Carlin C et al. (2019) Doorgym: a scalable door opening environment and baseline agent. arXiv preprint arXiv:1908.01887."},{"key":"e_1_3_5_63_1","unstructured":"Vaswani A Shazeer N Parmar N et al. (2017) Attention is all you need. In: Advances in Neural Information Processing Systems Long Beach December 4 to 9 2017."},{"key":"e_1_3_5_64_1","first-page":"2020","article-title":"Motion planner augmented reinforcement learning for obstructed environments. In: Conference on Robot Learning","volume":"16","author":"Yamada J","year":"2020","unstructured":"Yamada J, Lee Y, Salhotra G, et al. (2020) Motion planner augmented reinforcement learning for obstructed environments. In: Conference on Robot Learning. Virtual, November 16-18: 2020.","journal-title":"Virtual, November"},{"key":"e_1_3_5_65_1","doi-asserted-by":"crossref","unstructured":"Yang B Zhang J Pong V et al. (2019) Replab: a reproducible low-cost arm benchmark platform for robotic learning. arXiv preprint arXiv:1905.07447.","DOI":"10.1109\/ICRA.2019.8794390"},{"key":"e_1_3_5_66_1","unstructured":"Yu T Quillen D He Z et al (2019) Meta-world: a benchmark and evaluation for multi-task and meta reinforcement learning. In: Conference on Robot Learning Osaka Japan October 29 to 31 2019."},{"key":"e_1_3_5_67_1","doi-asserted-by":"publisher","DOI":"10.15607\/RSS.2023.XIX.016"},{"key":"e_1_3_5_68_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00589"},{"key":"e_1_3_5_69_1","unstructured":"Zhu Y Wong J Mandlekar A et al. (2020) robosuite: a modular simulation framework and benchmark for robot learning. arXiv preprint arXiv:2009.12293."},{"key":"e_1_3_5_70_1","unstructured":"Zitkovich B Yu T Xu S et al. (2023) RT-2: vision-language-action models transfer web knowledge to robotic control. In: Conference on Robot Learning. Atlanta Novemeber 6 to 9 2023. Available at: https:\/\/openreview.net\/forum?id=XMQgwiJ7KSX."}],"container-title":["The International Journal of Robotics Research"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/journals.sagepub.com\/doi\/pdf\/10.1177\/02783649241304789","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/journals.sagepub.com\/doi\/full-xml\/10.1177\/02783649241304789","content-type":"application\/xml","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/journals.sagepub.com\/doi\/pdf\/10.1177\/02783649241304789","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,4,29]],"date-time":"2026-04-29T10:17:24Z","timestamp":1777457844000},"score":1,"resource":{"primary":{"URL":"https:\/\/journals.sagepub.com\/doi\/10.1177\/02783649241304789"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,2,11]]},"references-count":69,"journal-issue":{"issue":"10-11","published-print":{"date-parts":[[2025,9]]}},"alternative-id":["10.1177\/02783649241304789"],"URL":"https:\/\/doi.org\/10.1177\/02783649241304789","relation":{},"ISSN":["0278-3649","1741-3176"],"issn-type":[{"value":"0278-3649","type":"print"},{"value":"1741-3176","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,2,11]]}}}