{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,29]],"date-time":"2026-05-29T11:13:05Z","timestamp":1780053185895,"version":"3.54.0"},"publisher-location":"New York, NY, USA","reference-count":115,"publisher":"ACM","content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2026,4,13]]},"DOI":"10.1145\/3772318.3791673","type":"proceedings-article","created":{"date-parts":[[2026,4,13]],"date-time":"2026-04-13T05:14:30Z","timestamp":1776057270000},"page":"1-23","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":4,"title":["Cocoa: Co-Planning and Co-Execution with AI Agents"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-2453-6315","authenticated-orcid":false,"given":"K. J. Kevin","family":"Feng","sequence":"first","affiliation":[{"name":"University of Washington, Seattle, Washington, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0000-4722-0631","authenticated-orcid":false,"given":"Kevin","family":"Pu","sequence":"additional","affiliation":[{"name":"University of Toronto, Toronto, Ontario, Canada"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-9984-1277","authenticated-orcid":false,"given":"Matt","family":"Latzke","sequence":"additional","affiliation":[{"name":"Allen Institute for AI, Seattle, Washington, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-6726-4009","authenticated-orcid":false,"given":"Tal","family":"August","sequence":"additional","affiliation":[{"name":"University of Illinois Urbana-Champaign, Urbana, Illinois, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0006-8042-885X","authenticated-orcid":false,"given":"Pao","family":"Siangliulue","sequence":"additional","affiliation":[{"name":"Allen Institute for AI, Seattle, Washington, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-5460-9047","authenticated-orcid":false,"given":"Jonathan","family":"Bragg","sequence":"additional","affiliation":[{"name":"Allen Institute for AI, Seattle, Washington, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-3255-0109","authenticated-orcid":false,"given":"Daniel S","family":"Weld","sequence":"additional","affiliation":[{"name":"Allen Institute for AI, Seattle, Washington, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-9462-9835","authenticated-orcid":false,"given":"Amy X.","family":"Zhang","sequence":"additional","affiliation":[{"name":"University of Washington, Seattle, Washington, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-0798-4351","authenticated-orcid":false,"given":"Joseph Chee","family":"Chang","sequence":"additional","affiliation":[{"name":"Allen Institute for AI, Seattle, Washington, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2026,4,13]]},"reference":[{"key":"e_1_3_3_2_2_2","doi-asserted-by":"crossref","unstructured":"Tyler Angert Miroslav Suzara Jenny Han Christopher Pondoc and Hariharan Subramonyam. 2023. Spellburst: A Node-based Interface for Exploratory Creative Coding with Natural Language Prompts. Proceedings of the 36th Annual ACM Symposium on User Interface Software and Technology (2023). https:\/\/api.semanticscholar.org\/CorpusID:260704716","DOI":"10.1145\/3586183.3606719"},{"key":"e_1_3_3_2_3_2","unstructured":"Anthropic. 2024. Developing a computer use model. https:\/\/www.anthropic.com\/news\/developing-computer-use."},{"key":"e_1_3_3_2_4_2","unstructured":"Ian Arawjo Chelse Swoopes Priyan Vaithilingam Martin Wattenberg and Elena\u00a0L. Glassman. 2023. ChainForge: A Visual Toolkit for Prompt Engineering and LLM Hypothesis Testing. ArXiv abs\/2309.09128 (2023). https:\/\/api.semanticscholar.org\/CorpusID:262044762"},{"key":"e_1_3_3_2_5_2","unstructured":"AutoGPT. 2024. Empower your digital tasks with AutoGPT. https:\/\/agpt.co\/."},{"key":"e_1_3_3_2_6_2","doi-asserted-by":"crossref","unstructured":"Amid Ayobi Jacob Hughes Christopher Duckworth Jakub\u00a0J Dylag Sam James Paul Marshall Matthew Guy Anitha Kumaran Adriane Chapman Michael\u00a0J. Boniface and Aisling\u00a0Ann O\u2019Kane. 2023. Computational Notebooks as Co-Design Tools: Engaging Young Adults Living with Diabetes Family Carers and Clinicians with Machine Learning Models. Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems (2023). https:\/\/api.semanticscholar.org\/CorpusID:258218057","DOI":"10.1145\/3544548.3581424"},{"key":"e_1_3_3_2_7_2","unstructured":"Jinheon Baek Sujay\u00a0Kumar Jauhar Silviu Cucerzan and Sung\u00a0Ju Hwang. 2024. ResearchAgent: Iterative Research Idea Generation over Scientific Literature with Large Language Models. ArXiv abs\/2404.07738 (2024). https:\/\/api.semanticscholar.org\/CorpusID:269042844"},{"key":"e_1_3_3_2_8_2","unstructured":"Gagan Bansal Jennifer\u00a0Wortman Vaughan Saleema Amershi Eric Horvitz Adam Fourney Hussein Mozannar Victor Dibia and Daniel\u00a0S Weld. 2024. Challenges in Human-Agent Communication. (2024). https:\/\/api.semanticscholar.org\/CorpusID:270870360"},{"key":"e_1_3_3_2_9_2","doi-asserted-by":"crossref","unstructured":"Dan Bennett Oussama Metatla Anne Roudaut and Elisa\u00a0D. Mekler. 2023. How does HCI Understand Human Agency and Autonomy? Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems (2023). https:\/\/api.semanticscholar.org\/CorpusID:256389761","DOI":"10.1145\/3544548.3580651"},{"key":"e_1_3_3_2_10_2","doi-asserted-by":"crossref","unstructured":"Virginia Braun and Victoria Clarke. 2019. Reflecting on reflexive thematic analysis. Qualitative Research in Sport Exercise and Health 11 (2019) 589 \u2013 597. https:\/\/api.semanticscholar.org\/CorpusID:197748828","DOI":"10.1080\/2159676X.2019.1628806"},{"key":"e_1_3_3_2_11_2","unstructured":"John Brooke et\u00a0al. 1996. SUS-A quick and dirty usability scale. Usability evaluation in industry 189 194 (1996) 4\u20137."},{"key":"e_1_3_3_2_12_2","doi-asserted-by":"publisher","DOI":"10.1145\/3630106.3658948"},{"key":"e_1_3_3_2_13_2","doi-asserted-by":"publisher","DOI":"10.1145\/3593013.3594033"},{"key":"e_1_3_3_2_14_2","doi-asserted-by":"crossref","unstructured":"Joel Chan Joseph\u00a0Chee Chang Tom Hope Dafna Shahaf and Aniket Kittur. 2018. Solvent: A mixed initiative system for finding analogies between research papers. Proceedings of the ACM on Human-Computer Interaction 2 CSCW (2018) 1\u201321.","DOI":"10.1145\/3274300"},{"key":"e_1_3_3_2_15_2","unstructured":"Joseph\u00a0Chee Chang Amy\u00a0X. Zhang Jonathan Bragg Andrew Head Kyle Lo Doug Downey and Daniel\u00a0S. Weld. 2023. CiteSee: Augmenting Citations in Scientific Papers with Persistent and Personalized Historical Context. Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems (2023). https:\/\/api.semanticscholar.org\/CorpusID:256868353"},{"key":"e_1_3_3_2_16_2","doi-asserted-by":"crossref","unstructured":"Souti Chattopadhyay I.\u00a0V. R. K.\u00a0V. Prasad Austin\u00a0Z. Henley Anita Sarma and Titus Barik. 2020. What\u2019s Wrong with Computational Notebooks? Pain Points Needs and Design Opportunities. Proceedings of the 2020 CHI Conference on Human Factors in Computing Systems (2020). https:\/\/api.semanticscholar.org\/CorpusID:210927488","DOI":"10.1145\/3313831.3376729"},{"key":"e_1_3_3_2_17_2","unstructured":"Jiangjie Chen Siyu Yuan Rong Ye Bodhisattwa\u00a0Prasad Majumder and Kyle Richardson. 2023. Put your money where your mouth is: Evaluating strategic planning and execution of llm agents in an auction arena. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2310.05746 (2023)."},{"key":"e_1_3_3_2_18_2","doi-asserted-by":"publisher","DOI":"10.1145\/3630106.3659048"},{"key":"e_1_3_3_2_19_2","unstructured":"Frederick Choi Sajjadur Rahman Han\u00a0Jun Kim and Daz Zhang. 2023. Towards Transparent Reusable and Customizable Data Science in Computational Notebooks. Extended Abstracts of the 2023 CHI Conference on Human Factors in Computing Systems (2023). https:\/\/api.semanticscholar.org\/CorpusID:257687372"},{"key":"e_1_3_3_2_20_2","unstructured":"Douglas\u00a0L. Dean Jillian\u00a0M. Hender Thomas\u00a0Lee Rodgers and Eric\u00a0L. Santanen. 2006. Identifying Quality Novel and Creative Ideas: Constructs and Scales for Idea Evaluation. J. Assoc. Inf. Syst. 7 (2006) 30. https:\/\/api.semanticscholar.org\/CorpusID:15910404"},{"key":"e_1_3_3_2_21_2","unstructured":"Elicit. 2024. Elicit: The AI Research Assistant. https:\/\/elicit.com\/."},{"key":"e_1_3_3_2_22_2","unstructured":"KJ Feng David\u00a0W McDonald and Amy\u00a0X Zhang. 2025. Levels of Autonomy for AI Agents. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2506.12469 (2025)."},{"key":"e_1_3_3_2_23_2","unstructured":"K.\u00a0J.\u00a0Kevin Feng Quan\u00a0Ze Chen Inyoung Cheong King Xia and Amy\u00a0X. Zhang. 2023. Case Repositories: Towards Case-Based Reasoning for AI Alignment. ArXiv abs\/2311.10934 (2023). https:\/\/api.semanticscholar.org\/CorpusID:265295304"},{"key":"e_1_3_3_2_24_2","doi-asserted-by":"crossref","unstructured":"Jennifer Fereday and Eimear Muir-Cochrane. 2006. Demonstrating rigor using thematic analysis: A hybrid approach of inductive and deductive coding and theme development. International journal of qualitative methods 5 1 (2006) 80\u201392.","DOI":"10.1177\/160940690600500107"},{"key":"e_1_3_3_2_25_2","unstructured":"Raymond Fok Joseph\u00a0Chee Chang Tal August Amy\u00a0X. Zhang and Daniel\u00a0S. Weld. 2023. Qlarify: Recursively Expandable Abstracts for Directed Information Retrieval over Scientific Papers. https:\/\/api.semanticscholar.org\/CorpusID:263835343"},{"key":"e_1_3_3_2_26_2","unstructured":"Raymond Fok Hita Kambhamettu Luca Soldaini Jonathan Bragg Kyle Lo Andrew Head Marti\u00a0A. Hearst and Daniel\u00a0S. Weld. 2022. Scim: Intelligent Skimming Support for Scientific Papers. Proceedings of the 28th International Conference on Intelligent User Interfaces (2022). https:\/\/api.semanticscholar.org\/CorpusID:254591867"},{"key":"e_1_3_3_2_27_2","unstructured":"Google. 2025. Gemini Deep Research. https:\/\/gemini.google\/overview\/deep-research\/."},{"key":"e_1_3_3_2_28_2","unstructured":"Juraj Gottweis Wei-Hung Weng Alexander Daryin Tao Tu Anil Palepu Petar Sirkovic Artiom Myaskovsky Felix Weissenberger Keran Rong Ryutaro Tanno et\u00a0al. 2025. Towards an AI co-scientist. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2502.18864 (2025)."},{"key":"e_1_3_3_2_29_2","unstructured":"Madeleine Grunde-McLaughlin Michelle\u00a0S. Lam Ranjay Krishna Daniel\u00a0S. Weld and Jeffrey Heer. 2023. Designing LLM Chains by Adapting Techniques from Crowdsourcing Workflows. ArXiv abs\/2312.11681 (2023). https:\/\/api.semanticscholar.org\/CorpusID:266362444"},{"key":"e_1_3_3_2_30_2","unstructured":"Joel Grus. 2018. I don\u2019t like notebooks. https:\/\/docs.google.com\/presentation\/d\/1n2RlMdmv1p25Xy5thJUhkKGvjtV-dkAIsUXP-AL4ffI\/edit#slide=id.g362da58057_0_1."},{"key":"e_1_3_3_2_31_2","unstructured":"Xuemei Gu and Mario Krenn. 2024. Generation and human-expert evaluation of interesting research ideas using knowledge graphs and large language models. https:\/\/api.semanticscholar.org\/CorpusID:270062620"},{"key":"e_1_3_3_2_32_2","unstructured":"Daya Guo Dejian Yang Haowei Zhang Junxiao Song Ruoyu Zhang Runxin Xu Qihao Zhu Shirong Ma Peiyi Wang Xiao Bi et\u00a0al. 2025. Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2501.12948 (2025)."},{"key":"e_1_3_3_2_33_2","doi-asserted-by":"crossref","unstructured":"Shibo Hao Yi Gu Haodi Ma Joshua\u00a0Jiahua Hong Zhen Wang Daisy\u00a0Zhe Wang and Zhiting Hu. 2023. Reasoning with Language Model is Planning with World Model. ArXiv abs\/2305.14992 (2023). https:\/\/api.semanticscholar.org\/CorpusID:258865812","DOI":"10.18653\/v1\/2023.emnlp-main.507"},{"key":"e_1_3_3_2_34_2","doi-asserted-by":"publisher","DOI":"10.1145\/3706598.3713218"},{"key":"e_1_3_3_2_35_2","doi-asserted-by":"publisher","DOI":"10.1145\/302979.303030"},{"key":"e_1_3_3_2_36_2","doi-asserted-by":"crossref","unstructured":"Faria Huq Zora\u00a0Zhiruo Wang Frank\u00a0F. Xu Tianyue Ou Shuyan Zhou Jeffrey\u00a0P. Bigham and Graham Neubig. 2025. CowPilot: A Framework for Autonomous and Human-Agent Collaborative Web Navigation. ArXiv abs\/2501.16609 (2025). https:\/\/api.semanticscholar.org\/CorpusID:275931872","DOI":"10.18653\/v1\/2025.naacl-demo.17"},{"key":"e_1_3_3_2_37_2","doi-asserted-by":"crossref","unstructured":"Edwin\u00a0L. Hutchins James Hollan and Donald\u00a0A. Norman. 1985. Direct Manipulation Interfaces. Hum. Comput. Interact. 1 (1985) 311\u2013338. https:\/\/api.semanticscholar.org\/CorpusID:16355120","DOI":"10.1207\/s15327051hci0104_2"},{"key":"e_1_3_3_2_38_2","unstructured":"Chip Huyen. 2025. Agents. https:\/\/huyenchip.com\/2025\/01\/07\/agents.html."},{"key":"e_1_3_3_2_39_2","unstructured":"Peter Jansen Marc-Alexandre C\u00f4t\u00e9 Tushar Khot Erin Bransom Bhavana\u00a0Dalvi Mishra Bodhisattwa\u00a0Prasad Majumder Oyvind Tafjord and Peter Clark. 2024. DISCOVERYWORLD: A Virtual Environment for Developing and Evaluating Automated Scientific Discovery Agents. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2406.06769 (2024)."},{"key":"e_1_3_3_2_40_2","doi-asserted-by":"crossref","unstructured":"Peter Jansen Oyvind Tafjord Marissa Radensky Pao Siangliulue Tom Hope Bhavana Dalvi Bodhisattwa\u00a0Prasad Majumder Daniel\u00a0S. Weld and Peter Clark. 2025. CodeScientist: End-to-End Semi-Automated Scientific Discovery with Code-based Experimentation. https:\/\/api.semanticscholar.org\/CorpusID:277451644","DOI":"10.18653\/v1\/2025.findings-acl.692"},{"key":"e_1_3_3_2_41_2","unstructured":"Carlos\u00a0E. Jimenez John Yang Alexander Wettig Shunyu Yao Kexin Pei Ofir Press and Karthik Narasimhan. 2023. SWE-bench: Can Language Models Resolve Real-World GitHub Issues? ArXiv abs\/2310.06770 (2023). https:\/\/api.semanticscholar.org\/CorpusID:263829697"},{"key":"e_1_3_3_2_42_2","doi-asserted-by":"crossref","unstructured":"Marina Jirotka Charlotte\u00a0P Lee and Gary\u00a0M Olson. 2013. Supporting scientific collaboration: Methods tools and concepts. Computer Supported Cooperative Work (CSCW) 22 (2013) 667\u2013715.","DOI":"10.1007\/s10606-012-9184-0"},{"key":"e_1_3_3_2_43_2","unstructured":"Project Jupyter. 2015. Jupyter Notebook UX Survey. https:\/\/github.com\/jupyter\/surveys\/tree\/master\/surveys\/2015-12-notebook-ux."},{"key":"e_1_3_3_2_44_2","unstructured":"Hyeonsu\u00a0B Kang Joseph\u00a0Chee Chang Yongsung Kim and Aniket Kittur. 2022. Threddy: An Interactive System for Personalized Thread-based Exploration and Organization of Scientific Literature. Proceedings of the 35th Annual ACM Symposium on User Interface Software and Technology (2022). https:\/\/api.semanticscholar.org\/CorpusID:251402552"},{"key":"e_1_3_3_2_45_2","unstructured":"Hyeonsu\u00a0B Kang Rafal Kocielnik Andrew Head Jiangjiang Yang Matt Latzke Aniket Kittur Daniel\u00a0S. Weld Doug Downey and Jonathan Bragg. 2022. From Who You Know to What You Read: Augmenting Scientific Recommendations with Implicit Social Networks. Proceedings of the 2022 CHI Conference on Human Factors in Computing Systems (2022). https:\/\/api.semanticscholar.org\/CorpusID:248299830"},{"key":"e_1_3_3_2_46_2","unstructured":"Hyeonsu\u00a0B Kang Sherry Wu Joseph\u00a0Chee Chang and Aniket Kittur. 2023. Synergi: A Mixed-Initiative System for Scholarly Synthesis and Sensemaking. Proceedings of the 36th Annual ACM Symposium on User Interface Software and Technology (2023). https:\/\/api.semanticscholar.org\/CorpusID:260899915"},{"key":"e_1_3_3_2_47_2","unstructured":"Sayash Kapoor Benedikt Stroebl Zachary\u00a0S. Siegel Nitya Nadgir and Arvind Narayanan. 2024. AI Agents That Matter. ArXiv abs\/2407.01502 (2024). https:\/\/api.semanticscholar.org\/CorpusID:270870360"},{"key":"e_1_3_3_2_48_2","doi-asserted-by":"crossref","unstructured":"Majeed Kazemitabaar Jack Williams Ian Drosos Tovi Grossman Austin\u00a0Z. Henley Carina Negreanu and Advait Sarkar. 2024. Improving Steering and Verification in AI-Assisted Data Analysis with Interactive Task Decomposition. https:\/\/api.semanticscholar.org\/CorpusID:270923956","DOI":"10.1145\/3654777.3676345"},{"key":"e_1_3_3_2_49_2","unstructured":"Mary\u00a0Beth Kery Marissa Radensky Mahima Arya Bonnie\u00a0E. John and Brad\u00a0A. Myers. 2018. The Story in the Notebook: Exploratory Data Science using a Literate Programming Tool. Proceedings of the 2018 CHI Conference on Human Factors in Computing Systems (2018). https:\/\/api.semanticscholar.org\/CorpusID:5060661"},{"key":"e_1_3_3_2_50_2","unstructured":"Mary\u00a0Beth Kery Donghao Ren Fred Hohman Dominik Moritz Kanit Wongsuphasawat and Kayur Patel. 2020. mage: Fluid Moves Between Code and Graphical Work in Computational Notebooks. Proceedings of the 33rd Annual ACM Symposium on User Interface Software and Technology (2020). https:\/\/api.semanticscholar.org\/CorpusID:221836345"},{"key":"e_1_3_3_2_51_2","unstructured":"Joongwon Kim Bhargavi Paranjape Tushar Khot and Hanna Hajishirzi. 2024. Husky: A Unified Open-Source Language Agent for Multi-Step Reasoning. https:\/\/api.semanticscholar.org\/CorpusID:270370824"},{"key":"e_1_3_3_2_52_2","doi-asserted-by":"publisher","DOI":"10.1145\/3563657.3595996"},{"key":"e_1_3_3_2_53_2","unstructured":"Rodney\u00a0Michael Kinney Chloe Anastasiades Russell Authur Iz Beltagy Jonathan Bragg Alexandra Buraczynski Isabel Cachola Stefan Candra Yoganand Chandrasekhar Arman Cohan Miles Crawford Doug Downey Jason Dunkelberger Oren Etzioni Rob Evans Sergey Feldman Joseph Gorney David\u00a0W. Graham F.Q. Hu Regan Huff Daniel King Sebastian Kohlmeier Bailey Kuehl Michael Langan Daniel Lin Haokun Liu Kyle Lo Jaron Lochner Kelsey MacMillan Tyler\u00a0C. Murray Christopher Newell Smita\u00a0R Rao Shaurya Rohatgi Paul Sayre Zejiang Shen Amanpreet Singh Luca Soldaini Shivashankar Subramanian A. Tanaka Alex\u00a0D Wade Linda\u00a0M. Wagner Lucy\u00a0Lu Wang Christopher Wilhelm Caroline Wu Jiangjiang Yang Angele Zamarron Madeleine van Zuylen and Daniel\u00a0S. Weld. 2023. The Semantic Scholar Open Data Platform. ArXiv abs\/2301.10140 (2023). https:\/\/api.semanticscholar.org\/CorpusID:256194545"},{"key":"e_1_3_3_2_54_2","doi-asserted-by":"publisher","DOI":"10.1093\/comjnl\/27.2.97"},{"key":"e_1_3_3_2_55_2","doi-asserted-by":"publisher","DOI":"10.1109\/VL\/HCC50065.2020.9127201"},{"key":"e_1_3_3_2_56_2","doi-asserted-by":"crossref","unstructured":"Lane Lawley and Christopher\u00a0J. MacLellan. 2024. VAL: Interactive Task Learning with GPT Dialog Parsing. Proceedings of the CHI Conference on Human Factors in Computing Systems (2024). https:\/\/api.semanticscholar.org\/CorpusID:263609076","DOI":"10.1145\/3613904.3641915"},{"key":"e_1_3_3_2_57_2","doi-asserted-by":"crossref","unstructured":"Charlotte\u00a0P Lee. 2007. Boundary negotiating artifacts: Unbinding the routine of boundary objects and embracing chaos in collaborative work. Computer Supported Cooperative Work (CSCW) 16 (2007) 307\u2013339.","DOI":"10.1007\/s10606-007-9044-5"},{"key":"e_1_3_3_2_58_2","doi-asserted-by":"crossref","unstructured":"Mina Lee Katy\u00a0Ilonka Gero John Joon\u00a0Young Chung Simon\u00a0Buckingham Shum Vipul Raheja Hua Shen Subhashini Venugopalan Thiemo Wambsganss David Zhou Emad\u00a0A. Alghamdi Tal August Avinash Bhat Madiha\u00a0Zahrah Choksi Senjuti Dutta Jin\u00a0L.C. Guo Md.\u00a0Naimul Hoque Yewon Kim Seyed\u00a0Parsa Neshaei Agnia Sergeyuk Antonette Shibani Disha Shrivastava Lila Shroff Jessi Stark S. Sterman Sitong Wang Antoine Bosselut Daniel Buschek Joseph\u00a0Chee Chang Sherol Chen Max Kreminski Joonsuk Park Roy Pea Eugenia\u00a0H. Rho Shannon\u00a0Zejiang Shen and Pao Siangliulue. 2024. A Design Space for Intelligent and Interactive Writing Assistants. Proceedings of the CHI Conference on Human Factors in Computing Systems (2024). https:\/\/api.semanticscholar.org\/CorpusID:268553985","DOI":"10.1145\/3613904.3642697"},{"key":"e_1_3_3_2_59_2","doi-asserted-by":"publisher","DOI":"10.1145\/3613904.3642196"},{"key":"e_1_3_3_2_60_2","doi-asserted-by":"publisher","DOI":"10.1145\/3544548.3580965"},{"key":"e_1_3_3_2_61_2","unstructured":"Zhehui Liao Maria Antoniak Inyoung Cheong Evie Yu-Yen Cheng Ai-Heng Lee Kyle Lo Joseph\u00a0Chee Chang and Amy\u00a0X Zhang. 2024. LLMs as Research Tools: A Large Scale Survey of Researchers\u2019 Usage and Perceptions. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2411.05025 (2024)."},{"key":"e_1_3_3_2_62_2","doi-asserted-by":"crossref","unstructured":"Henry Lieberman. 1997. Autonomous interface agents. Proceedings of the ACM SIGCHI Conference on Human factors in computing systems (1997). https:\/\/api.semanticscholar.org\/CorpusID:6576547","DOI":"10.1145\/258549.258592"},{"key":"e_1_3_3_2_63_2","unstructured":"Yanna Lin Haotian Li Leni Yang Aoyu Wu and Huamin Qu. 2023. InkSight: Leveraging Sketch Interaction for Documenting Chart Findings in Computational Notebooks. IEEE Transactions on Visualization and Computer Graphics 30 (2023) 944\u2013954. https:\/\/api.semanticscholar.org\/CorpusID:259936935"},{"key":"e_1_3_3_2_64_2","doi-asserted-by":"crossref","unstructured":"Yue Liu Sin\u00a0Kit Lo Qinghua Lu Liming Zhu Dehai Zhao Xiwei Xu Stefan Harrer and Jon Whittle. 2024. Agent Design Pattern Catalogue: A Collection of Architectural Patterns for Foundation Model based Agents. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2405.10467 (2024).","DOI":"10.2139\/ssrn.4922194"},{"key":"e_1_3_3_2_65_2","unstructured":"Chris Lu Cong Lu Robert\u00a0Tjarko Lange Jakob Foerster Jeff Clune and David Ha. 2024. The ai scientist: Towards fully automated open-ended scientific discovery. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2408.06292 (2024)."},{"key":"e_1_3_3_2_66_2","unstructured":"Xiao Ma Swaroop Mishra Ariel Liu Sophie\u00a0Ying Su Jilin Chen Chinmay Kulkarni Heng-Tze Cheng Quoc\u00a0V. Le and Ed\u00a0Huai hsin Chi. 2023. Beyond ChatBots: ExploreLLM for Structured Thoughts and Personalized Model Responses. Extended Abstracts of the CHI Conference on Human Factors in Computing Systems (2023). https:\/\/api.semanticscholar.org\/CorpusID:265551437"},{"key":"e_1_3_3_2_67_2","doi-asserted-by":"crossref","unstructured":"Pattie Maes. 1994. Agents that reduce work and information overload. Commun. ACM 37 (1994) 30\u201340. https:\/\/api.semanticscholar.org\/CorpusID:59868493","DOI":"10.1145\/176789.176792"},{"key":"e_1_3_3_2_68_2","doi-asserted-by":"crossref","unstructured":"Pattie Maes and Alan Wexelblat. 1996. Interface agents. Conference Companion on Human Factors in Computing Systems (1996). https:\/\/api.semanticscholar.org\/CorpusID:26827189","DOI":"10.1145\/257089.257377"},{"key":"e_1_3_3_2_69_2","unstructured":"Bodhisattwa\u00a0Prasad Majumder Harshit Surana Dhruv Agarwal Sanchaita Hazra Ashish Sabharwal and Peter Clark. 2024. Data-driven Discovery with Large Generative Models. ArXiv abs\/2402.13610 (2024). https:\/\/api.semanticscholar.org\/CorpusID:267770682"},{"key":"e_1_3_3_2_70_2","unstructured":"Hussein Mozannar Gagan Bansal Cheng Tan Adam Fourney Victor Dibia Jingya Chen Jack Gerrits Tyler Payne Matheus\u00a0Kunzler Maldaner Madeleine Grunde-McLaughlin Eric Zhu Griffin Bassman Jacob Alber Peter Chang Ricky Loynd Friederike Niedtner Ece Kamar Maya Murad Rafah Hosn and Saleema Amershi. 2025. Magentic-UI: Towards Human-in-the-loop Agentic Systems. ArXiv abs\/2507.22358 (2025). https:\/\/api.semanticscholar.org\/CorpusID:280391757"},{"key":"e_1_3_3_2_71_2","first-page":"7076","volume-title":"International conference on machine learning","author":"Mozannar Hussein","year":"2020","unstructured":"Hussein Mozannar and David Sontag. 2020. Consistent estimators for learning to defer to an expert. In International conference on machine learning. PMLR, 7076\u20137087."},{"key":"e_1_3_3_2_72_2","doi-asserted-by":"crossref","unstructured":"Niklas Muennighoff Zitong Yang Weijia Shi Xiang\u00a0Lisa Li Li Fei-Fei Hannaneh Hajishirzi Luke Zettlemoyer Percy Liang Emmanuel Cand\u00e8s and Tatsunori Hashimoto. 2025. s1: Simple test-time scaling. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2501.19393 (2025).","DOI":"10.18653\/v1\/2025.emnlp-main.1025"},{"key":"e_1_3_3_2_73_2","doi-asserted-by":"crossref","unstructured":"Jasper\u00a0Tran O\u2019Leary Gabrielle Benabdallah and Nadya Peek. 2023. Imprimer: Computational Notebooks for CNC Milling. Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems (2023). https:\/\/api.semanticscholar.org\/CorpusID:258217042","DOI":"10.1145\/3544548.3581334"},{"key":"e_1_3_3_2_74_2","unstructured":"OpenAI. 2024. Introducing canvas. https:\/\/openai.com\/index\/introducing-canvas\/."},{"key":"e_1_3_3_2_75_2","unstructured":"OpenAI. 2024. Learning to Reason with LLMs. https:\/\/openai.com\/index\/learning-to-reason-with-llms\/."},{"key":"e_1_3_3_2_76_2","unstructured":"OpenAI. 2025. ChatGPT Agent System Card. https:\/\/cdn.openai.com\/pdf\/839e66fc-602c-48bf-81d3-b21eacc3459d\/chatgpt_agent_system_card.pdf."},{"key":"e_1_3_3_2_77_2","unstructured":"OpenAI. 2025. Introducing deep research. https:\/\/openai.com\/index\/introducing-deep-research\/."},{"key":"e_1_3_3_2_78_2","unstructured":"OpenAI Platform. 2024. Assistants API. https:\/\/platform.openai.com\/docs\/assistants\/overview."},{"key":"e_1_3_3_2_79_2","doi-asserted-by":"crossref","unstructured":"James\u00a0C Overholser. 1993. Elements of the Socratic method: I. Systematic questioning. Psychotherapy: Theory Research Practice Training 30 1 (1993) 67.","DOI":"10.1037\/0033-3204.30.1.67"},{"key":"e_1_3_3_2_80_2","volume-title":"A keynote address delivered at the European Congress of Behavioural and Cognitive Therapies, London","author":"Padesky Christine\u00a0A","year":"1993","unstructured":"Christine\u00a0A Padesky. 1993. Socratic questioning: Changing minds or guiding discovery. In A keynote address delivered at the European Congress of Behavioural and Cognitive Therapies, London , Vol.\u00a024."},{"key":"e_1_3_3_2_81_2","doi-asserted-by":"crossref","unstructured":"Srishti Palani Aakanksha Naik Doug Downey Amy\u00a0X. Zhang Jonathan Bragg and Joseph\u00a0Chee Chang. 2023. Relatedly: Scaffolding Literature Reviews with Existing Related Work Sections. Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems (2023). https:\/\/api.semanticscholar.org\/CorpusID:256846632","DOI":"10.1145\/3544548.3580841"},{"key":"e_1_3_3_2_82_2","doi-asserted-by":"crossref","unstructured":"Yi-Hao Peng Dingzeyu Li Jeffrey\u00a0P. Bigham and Amy Pavel. 2025. Morae: Proactively Pausing UI Agents for User Choices. https:\/\/api.semanticscholar.org\/CorpusID:280985324","DOI":"10.1145\/3746059.3747797"},{"key":"e_1_3_3_2_83_2","unstructured":"Perplexity. 2024. Perplexity AI. https:\/\/www.perplexity.ai\/."},{"key":"e_1_3_3_2_84_2","unstructured":"Perplexity. 2025. Introducing Perplexity Deep Research. https:\/\/www.perplexity.ai\/hub\/blog\/introducing-perplexity-deep-research."},{"key":"e_1_3_3_2_85_2","unstructured":"Kevin Pu K.\u00a0J.\u00a0Kevin Feng Tovi Grossman Tom Hope Bhavana Dalvi Matt Latzke Jonathan Bragg Joseph\u00a0Chee Chang and Pao Siangliulue. 2024. IdeaSynth: Iterative Research Idea Development Through Evolving and Composing Idea Facets with Literature-Grounded Feedback. ArXiv abs\/2410.04025 (2024). https:\/\/api.semanticscholar.org\/CorpusID:273186404"},{"key":"e_1_3_3_2_86_2","unstructured":"Yujia Qin Shi Liang Yining Ye Kunlun Zhu Lan Yan Ya-Ting Lu Yankai Lin Xin Cong Xiangru Tang Bill Qian Sihan Zhao Runchu Tian Ruobing Xie Jie Zhou Marc\u00a0H. Gerstein Dahai Li Zhiyuan Liu and Maosong Sun. 2023. ToolLLM: Facilitating Large Language Models to Master 16000+ Real-world APIs. ArXiv abs\/2307.16789 (2023). https:\/\/api.semanticscholar.org\/CorpusID:260334759"},{"key":"e_1_3_3_2_87_2","doi-asserted-by":"crossref","unstructured":"Napol Rachatasumrit Jonathan Bragg Amy\u00a0X. Zhang and Daniel\u00a0S. Weld. 2022. CiteRead: Integrating Localized Citation Contexts into Scientific Paper Reading. 27th International Conference on Intelligent User Interfaces (2022). https:\/\/api.semanticscholar.org\/CorpusID:247585131","DOI":"10.1145\/3490099.3511162"},{"key":"e_1_3_3_2_88_2","unstructured":"Reworkd. 2024. AgentGPT. https:\/\/agentgpt.reworkd.ai\/."},{"key":"e_1_3_3_2_89_2","doi-asserted-by":"crossref","unstructured":"Adam Rule Aur\u00e9lien Tabard and James Hollan. 2018. Exploration and Explanation in Computational Notebooks. Proceedings of the 2018 CHI Conference on Human Factors in Computing Systems (2018). https:\/\/api.semanticscholar.org\/CorpusID:5048947","DOI":"10.1145\/3173574.3173606"},{"key":"e_1_3_3_2_90_2","doi-asserted-by":"publisher","DOI":"10.1145\/3672539.3695751"},{"key":"e_1_3_3_2_91_2","unstructured":"Timo Schick Jane Dwivedi-Yu Roberto Dess\u00ec Roberta Raileanu Maria Lomeli Luke Zettlemoyer Nicola Cancedda and Thomas Scialom. 2023. Toolformer: Language Models Can Teach Themselves to Use Tools. ArXiv abs\/2302.04761 (2023). https:\/\/api.semanticscholar.org\/CorpusID:256697342"},{"key":"e_1_3_3_2_92_2","doi-asserted-by":"crossref","unstructured":"Samuel Schmidgall Yusheng Su Ze Wang Ximeng Sun Jialian Wu Xiaodong Yu Jiang Liu Zicheng Liu and Emad Barsoum. 2025. Agent laboratory: Using llm agents as research assistants. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2501.04227 (2025).","DOI":"10.18653\/v1\/2025.findings-emnlp.320"},{"key":"e_1_3_3_2_93_2","unstructured":"Scite. 2025. AI for research. https:\/\/scite.ai\/."},{"key":"e_1_3_3_2_94_2","doi-asserted-by":"crossref","unstructured":"Omar Shaikh Shardul Sapkota Shan Rizvi Eric Horvitz Joon\u00a0Sung Park Diyi Yang and Michael\u00a0S Bernstein. 2025. Creating General User Models from Computer Use. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2505.10831 (2025).","DOI":"10.1145\/3746059.3747722"},{"key":"e_1_3_3_2_95_2","unstructured":"Yijia Shao Vinay Samuel Yucheng Jiang John Yang and Diyi Yang. 2024. Collaborative Gym: A Framework for Enabling and Evaluating Human-Agent Collaboration. https:\/\/api.semanticscholar.org\/CorpusID:274964990"},{"key":"e_1_3_3_2_96_2","unstructured":"Quan Shi Michael Tang Karthik Narasimhan and Shunyu Yao. 2024. Can Language Models Solve Olympiad Programming? ArXiv abs\/2404.10952 (2024). https:\/\/api.semanticscholar.org\/CorpusID:269187896"},{"key":"e_1_3_3_2_97_2","doi-asserted-by":"crossref","unstructured":"Ben Shneiderman and Pattie Maes. 1997. Direct manipulation vs. interface agents. Interactions 4 (1997) 42\u201361. https:\/\/api.semanticscholar.org\/CorpusID:27708923","DOI":"10.1145\/267505.267514"},{"key":"e_1_3_3_2_98_2","doi-asserted-by":"publisher","DOI":"10.7551\/mitpress\/7461.001.0001"},{"key":"e_1_3_3_2_99_2","doi-asserted-by":"crossref","unstructured":"Susan\u00a0Leigh Star and James\u00a0R Griesemer. 1989. Institutional ecology translations\u2019 and boundary objects: Amateurs and professionals in Berkeley\u2019s Museum of Vertebrate Zoology 1907-39. Social studies of science 19 3 (1989) 387\u2013420.","DOI":"10.1177\/030631289019003001"},{"key":"e_1_3_3_2_100_2","unstructured":"Sangho Suh Bryan Min Srishti Palani and Haijun Xia. 2023. Sensecape: Enabling Multilevel Exploration and Sensemaking with Large Language Models. Proceedings of the 36th Annual ACM Symposium on User Interface Software and Technology (2023). https:\/\/api.semanticscholar.org\/CorpusID:258822925"},{"key":"e_1_3_3_2_101_2","unstructured":"Theodore\u00a0R. Sumers Shunyu Yao Karthik Narasimhan and Thomas\u00a0L. Griffiths. 2023. Cognitive Architectures for Language Agents. ArXiv abs\/2309.02427 (2023). https:\/\/api.semanticscholar.org\/CorpusID:261556862"},{"key":"e_1_3_3_2_102_2","unstructured":"Karthik Valmeekam Matthew Marquez Sarath Sreedharan and Subbarao Kambhampati. 2023. On the Planning Abilities of Large Language Models - A Critical Investigation. ArXiv abs\/2305.15771 (2023). https:\/\/api.semanticscholar.org\/CorpusID:260440590"},{"key":"e_1_3_3_2_103_2","unstructured":"Guanzhi Wang Yuqi Xie Yunfan Jiang Ajay Mandlekar Chaowei Xiao Yuke Zhu Linxi\u00a0(Jim) Fan and Anima Anandkumar. 2023. Voyager: An Open-Ended Embodied Agent with Large Language Models. ArXiv abs\/2305.16291 (2023). https:\/\/api.semanticscholar.org\/CorpusID:258887849"},{"key":"e_1_3_3_2_104_2","doi-asserted-by":"crossref","unstructured":"Xingbo Wang Samantha\u00a0Lee Huey Rui Sheng Saurabh Mehta and Fei Wang. 2024. SciDaSynth: Interactive Structured Knowledge Extraction and Synthesis from Scientific Literature with Large Language Model. ArXiv abs\/2404.13765 (2024). https:\/\/api.semanticscholar.org\/CorpusID:269293213","DOI":"10.1002\/CL2.70073\/v2\/response1"},{"key":"e_1_3_3_2_105_2","unstructured":"Jason Wei Xuezhi Wang Dale Schuurmans Maarten Bosma Ed\u00a0Huai hsin Chi F. Xia Quoc Le and Denny Zhou. 2022. Chain of Thought Prompting Elicits Reasoning in Large Language Models. ArXiv abs\/2201.11903 (2022). https:\/\/api.semanticscholar.org\/CorpusID:246411621"},{"key":"e_1_3_3_2_106_2","unstructured":"Scott Wu. 2024. Introducing Devin the first AI software engineer. https:\/\/www.cognition.ai\/blog\/introducing-devin."},{"key":"e_1_3_3_2_107_2","unstructured":"Tongshuang\u00a0Sherry Wu Haiyi Zhu Maya Albayrak Alexis Axon Amanda Bertsch Wenxing Deng Ziqi Ding Bill\u00a0Boyuan Guo Sireesh Gururaja Tzu-Sheng Kuo Jenny\u00a0T Liang Ryan Liu Ihita Mandal Jeremiah Milbauer Xiaolin Ni N. Padmanabhan Subhashini Ramkumar Alexis Sudjianto Jordan Taylor Ying-Jui Tseng Patricia Vaidos Zhijin Wu Wei Wu and Chenyang Yang. 2023. LLMs as Workers in Human-Computational Algorithms? Replicating Crowdsourcing Pipelines with LLMs. ArXiv abs\/2307.10168 (2023). https:\/\/api.semanticscholar.org\/CorpusID:259982473"},{"key":"e_1_3_3_2_108_2","unstructured":"John Yang Carlos\u00a0E. Jimenez Alexander Wettig Kilian Lieret Shunyu Yao Karthik Narasimhan and Ofir Press. 2024. SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering. ArXiv abs\/2405.15793 (2024). https:\/\/api.semanticscholar.org\/CorpusID:270063685"},{"key":"e_1_3_3_2_109_2","doi-asserted-by":"crossref","unstructured":"Shunyu Yao Howard Chen John Yang and Karthik Narasimhan. 2022. Webshop: Towards scalable real-world web interaction with grounded language agents. Advances in Neural Information Processing Systems 35 (2022) 20744\u201320757.","DOI":"10.52202\/068431-1508"},{"key":"e_1_3_3_2_110_2","unstructured":"Shunyu Yao Dian Yu Jeffrey Zhao Izhak Shafran Thomas\u00a0L. Griffiths Yuan Cao and Karthik Narasimhan. 2023. Tree of Thoughts: Deliberate Problem Solving with Large Language Models. ArXiv abs\/2305.10601 (2023). https:\/\/api.semanticscholar.org\/CorpusID:258762525"},{"key":"e_1_3_3_2_111_2","unstructured":"Shunyu Yao Jeffrey Zhao Dian Yu Nan Du Izhak Shafran Karthik Narasimhan and Yuan Cao. 2022. ReAct: Synergizing Reasoning and Acting in Language Models. ArXiv abs\/2210.03629 (2022). https:\/\/api.semanticscholar.org\/CorpusID:252762395"},{"key":"e_1_3_3_2_112_2","doi-asserted-by":"crossref","unstructured":"Xingjian Zhang Yutong Xie Jin Huang Jinge Ma Zhaoying Pan Qijia Liu Ziyang Xiong Tolga Ergen Dongsub Shim Honglak Lee and Qiaozhu Mei. 2024. MASSW: A New Dataset and Benchmark Tasks for AI-Assisted Scientific Workflows. https:\/\/api.semanticscholar.org\/CorpusID:270370776","DOI":"10.18653\/v1\/2025.findings-naacl.127"},{"key":"e_1_3_3_2_113_2","doi-asserted-by":"publisher","DOI":"10.1145\/3586183.3606800"},{"key":"e_1_3_3_2_114_2","doi-asserted-by":"crossref","unstructured":"Chengbo Zheng Dakuo Wang April\u00a0Yi Wang and Xiaojuan Ma. 2022. Telling Stories from Computational Notebooks: AI-Assisted Presentation Slides Creation for Presenting Data Science Work. Proceedings of the 2022 CHI Conference on Human Factors in Computing Systems (2022). https:\/\/api.semanticscholar.org\/CorpusID:247594488","DOI":"10.1145\/3491102.3517615"},{"key":"e_1_3_3_2_115_2","unstructured":"Andy Zhou Kai Yan Michal Shlapentokh-Rothman Haohan Wang and Yu-Xiong Wang. 2023. Language Agent Tree Search Unifies Reasoning Acting and Planning in Language Models. ArXiv abs\/2310.04406 (2023). https:\/\/api.semanticscholar.org\/CorpusID:263829963"},{"key":"e_1_3_3_2_116_2","unstructured":"Yuchen Zhuang Xiang Chen Tong Yu Saayan Mitra Victor\u00a0S. Bursztyn Ryan\u00a0A. Rossi Somdeb Sarkhel and Chao Zhang. 2023. ToolChain*: Efficient Action Space Navigation in Large Language Models with A* Search. ArXiv abs\/2310.13227 (2023). https:\/\/api.semanticscholar.org\/CorpusID:264405734"}],"event":{"name":"CHI 2026: CHI Conference on Human Factors in Computing Systems","location":"Barcelona Spain","acronym":"CHI '26","sponsor":["SIGCHI ACM Special Interest Group on Computer-Human Interaction"]},"container-title":["Proceedings of the 2026 CHI Conference on Human Factors in Computing Systems"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3772318.3791673","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,4,17]],"date-time":"2026-04-17T09:19:22Z","timestamp":1776417562000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3772318.3791673"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,4,13]]},"references-count":115,"alternative-id":["10.1145\/3772318.3791673","10.1145\/3772318"],"URL":"https:\/\/doi.org\/10.1145\/3772318.3791673","relation":{},"subject":[],"published":{"date-parts":[[2026,4,13]]},"assertion":[{"value":"2026-04-13","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}