{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,22]],"date-time":"2026-05-22T04:07:45Z","timestamp":1779422865239,"version":"3.53.1"},"publisher-location":"New York, NY, USA","reference-count":64,"publisher":"ACM","license":[{"start":{"date-parts":[[2026,5,26]],"date-time":"2026-05-26T00:00:00Z","timestamp":1779753600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/legalcode"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2026,5,26]]},"DOI":"10.1145\/3786335.3813168","type":"proceedings-article","created":{"date-parts":[[2026,5,22]],"date-time":"2026-05-22T03:16:22Z","timestamp":1779419782000},"page":"917-952","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["Scaling Textual Gradients via Sampling-Based Momentum"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0009-0007-2227-1046","authenticated-orcid":false,"given":"Zixin","family":"Ding","sequence":"first","affiliation":[{"name":"University of Chicago, Chicago, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-5718-5187","authenticated-orcid":false,"given":"Junyuan","family":"Hong","sequence":"additional","affiliation":[{"name":"University of Texas at Austin, Austin, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0002-1795-3409","authenticated-orcid":false,"given":"Zhan","family":"Shi","sequence":"additional","affiliation":[{"name":"Santa Clara University, Santa Clara, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-7944-3418","authenticated-orcid":false,"given":"Tianhao","family":"Wang","sequence":"additional","affiliation":[{"name":"Princeton University, Princeton, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-8421-2662","authenticated-orcid":false,"given":"Zinan","family":"Lin","sequence":"additional","affiliation":[{"name":"Microsoft Research, Redmond, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-1063-0682","authenticated-orcid":false,"given":"Li","family":"Yin","sequence":"additional","affiliation":[{"name":"SylphAI, San Francisco, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-9420-3874","authenticated-orcid":false,"given":"Meng","family":"Liu","sequence":"additional","affiliation":[{"name":"SylphAI, San Francisco, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-2050-5693","authenticated-orcid":false,"given":"Zhangyang","family":"Wang","sequence":"additional","affiliation":[{"name":"University of Texas at Austin, Austin, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-2133-140X","authenticated-orcid":false,"given":"Yuxin","family":"Chen","sequence":"additional","affiliation":[{"name":"University of Chicago, Chicago, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2026,5,26]]},"reference":[{"key":"e_1_3_3_1_2_2","doi-asserted-by":"crossref","unstructured":"Rishabh Agarwal Avi Singh Lei Zhang Bernd Bohnet Luis Rosias Stephanie Chan Biao Zhang Ankesh Anand Zaheer Abbas Azade Nova et\u00a0al. 2024. Many-shot in-context learning. Advances in Neural Information Processing Systems 37 (2024) 76930\u201376966.","DOI":"10.52202\/079017-2447"},{"key":"e_1_3_3_1_3_2","unstructured":"Lakshya\u00a0A Agrawal Shangyin Tan Dilara Soylu Noah Ziems Rishi Khare Krista Opsahl-Ong Arnav Singhvi Herumb Shandilya Michael\u00a0J Ryan Meng Jiang et\u00a0al. 2025. Gepa: Reflective prompt evolution can outperform reinforcement learning. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2507.19457 (2025)."},{"key":"e_1_3_3_1_4_2","volume-title":"COLM","author":"Anagnostidis Sotiris","year":"2024","unstructured":"Sotiris Anagnostidis and Jannis Bulian. 2024. How Susceptible are LLMs to Influence in Prompts?. In COLM."},{"key":"e_1_3_3_1_5_2","doi-asserted-by":"crossref","unstructured":"Leo Breiman. 1996. Bagging predictors. Machine learning 24 2 (1996) 123\u2013140.","DOI":"10.1023\/A:1018054314350"},{"key":"e_1_3_3_1_6_2","unstructured":"Bradley Brown Jordan Juravsky Ryan Ehrlich Ronald Clark Quoc\u00a0V Le Christopher R\u00e9 and Azalia Mirhoseini. 2024. Large language monkeys: Scaling inference compute with repeated sampling. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2407.21787 (2024)."},{"key":"e_1_3_3_1_7_2","unstructured":"Tom Brown Benjamin Mann Nick Ryder Melanie Subbiah Jared\u00a0D Kaplan Prafulla Dhariwal Arvind Neelakantan Pranav Shyam Girish Sastry Amanda Askell et\u00a0al. 2020. Language models are few-shot learners. Advances in neural information processing systems 33 (2020) 1877\u20131901."},{"key":"e_1_3_3_1_8_2","volume-title":"Forty-second International Conference on Machine Learning","author":"Chen Yaofo","year":"2024","unstructured":"Yaofo Chen, Zeng You, Shuhai Zhang, Haokun Li, Yirui Li, Yaowei Wang, and Mingkui Tan. 2024. Core Context Aware Transformers for Long Context Language Modeling. In Forty-second International Conference on Machine Learning."},{"key":"e_1_3_3_1_9_2","unstructured":"Peter Clark Isaac Cowhey Oren Etzioni Tushar Khot Ashish Sabharwal Carissa Schoenick and Oyvind Tafjord. 2018. Think you have solved question answering? try arc the ai2 reasoning challenge. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/1803.05457 (2018)."},{"key":"e_1_3_3_1_10_2","unstructured":"Karl Cobbe Vineet Kosaraju Mohammad Bavarian Mark Chen Heewoo Jun Lukasz Kaiser Matthias Plappert Jerry Tworek Jacob Hilton Reiichiro Nakano et\u00a0al. 2021. Training verifiers to solve math word problems. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2110.14168 (2021)."},{"key":"e_1_3_3_1_11_2","unstructured":"Gheorghe Comanici Eric Bieber Mike Schaekermann Ice Pasupat Noveen Sachdeva Inderjit Dhillon Marcel Blistein Ori Ram Dan Zhang Evan Rosen et\u00a0al. 2025. Gemini 2.5: Pushing the frontier with advanced reasoning multimodality long context and next generation agentic capabilities. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2507.06261 (2025)."},{"key":"e_1_3_3_1_12_2","volume-title":"The Fourteenth International Conference on Learning Representations","author":"Llano Enrique\u00a0Queipo de","year":"2026","unstructured":"Enrique\u00a0Queipo de Llano, Alvaro Arroyo, Federico Barbero, Xiaowen Dong, Michael\u00a0M. Bronstein, Yann LeCun, and Ravid Shwartz-Ziv. 2026. Attention Sinks and Compression Valleys in LLMs are Two Sides of the Same Coin. In The Fourteenth International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=c5TFhCJ6fs"},{"key":"e_1_3_3_1_13_2","unstructured":"DSPy. 2025. Math Reasoning \u2013 DSPy Tutorial. https:\/\/dspy.ai\/tutorials\/math\/. Accessed: 2025-09-18."},{"key":"e_1_3_3_1_14_2","unstructured":"DSPy. 2025. Optimization \/ Optimizers. https:\/\/dspy.ai\/learn\/optimization\/optimizers\/. Accessed: 2025-05-14."},{"key":"e_1_3_3_1_15_2","unstructured":"DSPy Team. 2025. Built in datasets. https:\/\/dspy.ai\/deep-dive\/data-handling\/built-in-datasets\/. DSPy Documentation. Accessed 2025-10-11."},{"key":"e_1_3_3_1_16_2","unstructured":"DSPy Team. 2025. Tutorial: Math Reasoning. https:\/\/dspy.ai\/tutorials\/math\/. Accessed: 2025-10-12."},{"key":"e_1_3_3_1_17_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2025.findings-emnlp.1264"},{"key":"e_1_3_3_1_18_2","unstructured":"Yaru Hao Yutao Sun Li Dong Zhixiong Han Yuxian Gu and Furu Wei. 2022. Structured prompting: Scaling in-context learning to 1 000 examples. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2212.06713 (2022)."},{"key":"e_1_3_3_1_19_2","volume-title":"Thirty-fifth Conference on Neural Information Processing Systems Datasets and Benchmarks Track (Round 2)","author":"Hendrycks Dan","year":"2021","unstructured":"Dan Hendrycks, Collin Burns, Saurav Kadavath, Akul Arora, Steven Basart, Eric Tang, Dawn Song, and Jacob Steinhardt. 2021. Measuring Mathematical Problem Solving With the MATH Dataset. In Thirty-fifth Conference on Neural Information Processing Systems Datasets and Benchmarks Track (Round 2)."},{"key":"e_1_3_3_1_20_2","unstructured":"Jordan Hoffmann Sebastian Borgeaud Arthur Mensch Elena Buchatskaya Trevor Cai Eliza Rutherford Diego de\u00a0Las Casas Lisa\u00a0Anne Hendricks Johannes Welbl Aidan Clark et\u00a0al. 2022. Training compute-optimal large language models. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2203.15556 (2022)."},{"key":"e_1_3_3_1_21_2","unstructured":"Coleman Hooper Sehoon Kim Hiva Mohammadzadeh Monishwaran Maheswaran June Paik Michael\u00a0W Mahoney Kurt Keutzer and Amir Gholami. 2024. Squeezed attention: Accelerating long context length llm inference. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2411.09688 (2024)."},{"key":"e_1_3_3_1_22_2","unstructured":"Jared Kaplan Sam McCandlish Tom Henighan Tom\u00a0B Brown Benjamin Chess Rewon Child Scott Gray Alec Radford Jeffrey Wu and Dario Amodei. 2020. Scaling laws for neural language models. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2001.08361 (2020)."},{"key":"e_1_3_3_1_23_2","volume-title":"The Twelfth International Conference on Learning Representations","author":"Khattab Omar","year":"2024","unstructured":"Omar Khattab, Arnav Singhvi, Paridhi Maheshwari, Zhiyuan Zhang, Keshav Santhanam, Saiful Haq, Ashutosh Sharma, Thomas\u00a0T Joshi, Hanna Moazam, Heather Miller, et\u00a0al. 2024. Dspy: Compiling declarative language model calls into state-of-the-art pipelines. In The Twelfth International Conference on Learning Representations."},{"key":"e_1_3_3_1_24_2","unstructured":"Andreas Kirsch Sebastian Farquhar Parmida Atighehchian Andrew Jesson Fr\u00e9d\u00e9ric Branchaud-Charron and Yarin Gal. 2023. Stochastic Batch Acquisition: A Simple Baseline for Deep Active Learning. Transactions on Machine Learning Research (2023)."},{"key":"e_1_3_3_1_25_2","first-page":"3499","volume-title":"International conference on machine learning","author":"Kool Wouter","year":"2019","unstructured":"Wouter Kool, Herke Van\u00a0Hoof, and Max Welling. 2019. Stochastic beams and where to find them: The gumbel-top-k trick for sampling sequences without replacement. In International conference on machine learning. PMLR, 3499\u20133508."},{"key":"e_1_3_3_1_26_2","doi-asserted-by":"crossref","unstructured":"Yury Kuratov Aydar Bulatov Petr Anokhin Ivan Rodkin Dmitry Sorokin Artyom Sorokin and Mikhail Burtsev. 2024. Babilong: Testing the limits of llms with long context reasoning-in-a-haystack. Advances in Neural Information Processing Systems 37 (2024) 106519\u2013106554.","DOI":"10.52202\/079017-3381"},{"key":"e_1_3_3_1_27_2","doi-asserted-by":"publisher","DOI":"10.1017\/9781108571401"},{"key":"e_1_3_3_1_28_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2024.acl-long.818"},{"key":"e_1_3_3_1_29_2","unstructured":"Dacheng Li Shiyi Cao Chengkun Cao Xiuyu Li Shangyin Tan Kurt Keutzer Jiarong Xing Joseph\u00a0E Gonzalez and Ion Stoica. 2025. S*: Test time scaling for code generation. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2502.14382 (2025)."},{"key":"e_1_3_3_1_30_2","doi-asserted-by":"crossref","unstructured":"Nelson\u00a0F Liu Kevin Lin John Hewitt Ashwin Paranjape Michele Bevilacqua Fabio Petroni and Percy Liang. 2024. Lost in the middle: How language models use long contexts. Transactions of the Association for Computational Linguistics 12 (2024) 157\u2013173.","DOI":"10.1162\/tacl_a_00638"},{"key":"e_1_3_3_1_31_2","unstructured":"Xiaoxuan Liu Cade Daniel Langxiang Hu Woosuk Kwon Zhuohan Li Xiangxi Mo Alvin Cheung Zhijie Deng Ion Stoica and Hao Zhang. 2024. Optimizing speculative decoding for serving large language models using goodput. arXiv e-prints (2024) arXiv\u20132406."},{"key":"e_1_3_3_1_32_2","unstructured":"Yanli Liu Yuan Gao and Wotao Yin. 2020. An improved analysis of stochastic gradient descent with momentum. Advances in Neural Information Processing Systems 33 (2020) 18261\u201318271."},{"key":"e_1_3_3_1_33_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2022.acl-long.556"},{"key":"e_1_3_3_1_34_2","unstructured":"Chris\u00a0J Maddison Daniel Tarlow and Tom Minka. 2014. A* sampling. Advances in neural information processing systems 27 (2014)."},{"key":"e_1_3_3_1_35_2","doi-asserted-by":"crossref","unstructured":"Niklas Muennighoff Zitong Yang Weijia Shi Xiang\u00a0Lisa Li Li Fei-Fei Hannaneh Hajishirzi Luke Zettlemoyer Percy Liang Emmanuel Cand\u00e8s and Tatsunori Hashimoto. 2025. s1: Simple test-time scaling. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2501.19393 (2025).","DOI":"10.18653\/v1\/2025.emnlp-main.1025"},{"key":"e_1_3_3_1_36_2","doi-asserted-by":"crossref","unstructured":"Xuefei Ning Zifu Wang Shiyao Li Zinan Lin Peiran Yao Tianyu Fu Matthew Blaschko Guohao Dai Huazhong Yang and Yu Wang. 2024. Can LLMs learn by teaching for better reasoning? A preliminary study. Advances in Neural Information Processing Systems 37 (2024) 71188\u201371239.","DOI":"10.52202\/079017-2275"},{"key":"e_1_3_3_1_37_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2024.emnlp-main.525"},{"key":"e_1_3_3_1_38_2","doi-asserted-by":"crossref","unstructured":"Long Ouyang Jeffrey Wu Xu Jiang Diogo Almeida Carroll Wainwright Pamela Mishkin Chong Zhang Sandhini Agarwal Katarina Slama Alex Ray et\u00a0al. 2022. Training language models to follow instructions with human feedback. Advances in neural information processing systems 35 (2022) 27730\u201327744.","DOI":"10.52202\/068431-2011"},{"key":"e_1_3_3_1_39_2","unstructured":"Hao Peng Xiaozhi Wang Jianhui Chen Weikai Li Yunjia Qi Zimu Wang Zhili Wu Kaisheng Zeng Bin Xu Lei Hou et\u00a0al. 2023. When does in-context learning fall short and why? a study on specification-heavy tasks. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2311.08993 (2023)."},{"key":"e_1_3_3_1_40_2","doi-asserted-by":"crossref","unstructured":"Boris\u00a0T Polyak. 1964. Some methods of speeding up the convergence of iteration methods. Ussr computational mathematics and mathematical physics 4 5 (1964) 1\u201317.","DOI":"10.1016\/0041-5553(64)90137-5"},{"key":"e_1_3_3_1_41_2","doi-asserted-by":"crossref","unstructured":"Reid Pryzant Dan Iter Jerry Li Yin\u00a0Tat Lee Chenguang Zhu and Michael Zeng. 2023. Automatic prompt optimization with\" gradient descent\" and beam search. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2305.03495 (2023).","DOI":"10.18653\/v1\/2023.emnlp-main.494"},{"key":"e_1_3_3_1_42_2","volume-title":"The Thirty-ninth Annual Conference on Neural Information Processing Systems Datasets and Benchmarks Track","author":"Pyatkin Valentina","year":"2024","unstructured":"Valentina Pyatkin, Saumya Malik, Victoria Graf, Hamish Ivison, Shengyi Huang, Pradeep Dasigi, Nathan Lambert, and Hannaneh Hajishirzi. 2024. Generalizing Verifiable Instruction Following. In The Thirty-ninth Annual Conference on Neural Information Processing Systems Datasets and Benchmarks Track."},{"key":"e_1_3_3_1_43_2","unstructured":"Prithvi Rajasekaran Ethan Dixon Carly Ryan and Jeremy Hadfield. 2025. Effective Context Engineering for AI Agents. Anthropic Engineering Blog. https:\/\/www.anthropic.com\/engineering\/effective-context-engineering-for-ai-agents With contributions from Rafi Ayub Hannah Moran Cal Rueb and Connor Jennings."},{"key":"e_1_3_3_1_44_2","doi-asserted-by":"crossref","unstructured":"David\u00a0E Rumelhart Geoffrey\u00a0E Hinton and Ronald\u00a0J Williams. 1986. Learning representations by back-propagating errors. nature 323 6088 (1986) 533\u2013536.","DOI":"10.1038\/323533a0"},{"key":"e_1_3_3_1_45_2","volume-title":"The Twelfth International Conference on Learning Representations","author":"Sclar Melanie","year":"2024","unstructured":"Melanie Sclar, Yejin Choi, Yulia Tsvetkov, and Alane Suhr. 2024. Quantifying Language Models\u2019 Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formatting. In The Twelfth International Conference on Learning Representations."},{"key":"e_1_3_3_1_46_2","unstructured":"Charlie Snell Jaehoon Lee Kelvin Xu and Aviral Kumar. 2024. Scaling llm test-time compute optimally can be more effective than scaling model parameters. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2408.03314 (2024)."},{"key":"e_1_3_3_1_47_2","doi-asserted-by":"crossref","unstructured":"Alessandro Sordoni Eric Yuan Marc-Alexandre C\u00f4t\u00e9 Matheus Pereira Adam Trischler Ziang Xiao Arian Hosseini Friederike Niedtner and Nicolas Le\u00a0Roux. 2023. Joint prompt optimization of stacked llms using variational inference. Advances in Neural Information Processing Systems 36 (2023) 58128\u201358151.","DOI":"10.52202\/075280-2534"},{"key":"e_1_3_3_1_48_2","unstructured":"Jiaming Tang Yilong Zhao Kan Zhu Guangxuan Xiao Baris Kasikci and Song Han. 2024. Quest: Query-aware sparsity for efficient long-context llm inference. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2406.10774 (2024)."},{"key":"e_1_3_3_1_49_2","unstructured":"Thinking Machines. 2025. Defeating Nondeterminism in LLM Inference. https:\/\/thinkingmachines.ai\/blog\/defeating-nondeterminism-in-llm-inference\/"},{"key":"e_1_3_3_1_50_2","volume-title":"Advances in Neural Information Processing Systems","author":"Vaswani Ashish","year":"2017","unstructured":"Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan\u00a0N Gomez, \u0141\u00a0ukasz Kaiser, and Illia Polosukhin. 2017. Attention is All you Need. In Advances in Neural Information Processing Systems, Vol.\u00a030. Curran Associates, Inc.https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2017\/file\/3f5ee243547dee91fbd053c1c4a845aa-Paper.pdf"},{"key":"e_1_3_3_1_51_2","unstructured":"vLLM Project. 2024. vLLM Documentation: Automatic Prefix Caching. https:\/\/docs.vllm.ai\/en\/latest\/design\/prefix_caching\/. Accessed 2026-02-18."},{"key":"e_1_3_3_1_52_2","doi-asserted-by":"crossref","unstructured":"Xingchen Wan Ruoxi Sun Hootan Nakhost and Sercan Arik. 2024. Teach better or show smarter? on instructions and exemplars in automatic prompt optimization. Advances in Neural Information Processing Systems 37 (2024) 58174\u201358244.","DOI":"10.52202\/079017-1855"},{"key":"e_1_3_3_1_53_2","volume-title":"The Twelfth International Conference on Learning Representations","author":"Wang Xinyuan","year":"2024","unstructured":"Xinyuan Wang, Chenxi Li, Zhen Wang, Fan Bai, Haotian Luo, Jiayou Zhang, Nebojsa Jojic, Eric Xing, and Zhiting Hu. 2024. PromptAgent: Strategic Planning with Language Models Enables Expert-level Prompt Optimization. In The Twelfth International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=22pyNMuIoa"},{"key":"e_1_3_3_1_54_2","volume-title":"The Twelfth International Conference on Learning Representations","author":"Xiao Guangxuan","year":"2024","unstructured":"Guangxuan Xiao, Yuandong Tian, Beidi Chen, Song Han, and Mike Lewis. 2024. Efficient Streaming Language Models with Attention Sinks. In The Twelfth International Conference on Learning Representations."},{"key":"e_1_3_3_1_55_2","volume-title":"The Twelfth International Conference on Learning Representations","author":"Yang Chengrun","year":"2024","unstructured":"Chengrun Yang, Xuezhi Wang, Yifeng Lu, Hanxiao Liu, Quoc\u00a0V Le, Denny Zhou, and Xinyun Chen. 2024. Large Language Models as Optimizers. In The Twelfth International Conference on Learning Representations."},{"key":"e_1_3_3_1_56_2","unstructured":"Yunhao Yang Junyuan Hong Gabriel\u00a0Jacob Perin Zhiwen Fan Li Yin Zhangyang Wang and Ufuk Topcu. 2025. AD-VF: LLM-Automatic Differentiation Enables Fine-Tuning-Free Robot Planning from Formal Methods Feedback. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2509.18384 (2025)."},{"key":"e_1_3_3_1_57_2","doi-asserted-by":"crossref","unstructured":"Zhilin Yang Peng Qi Saizheng Zhang Yoshua Bengio William\u00a0W Cohen Ruslan Salakhutdinov and Christopher\u00a0D Manning. 2018. HotpotQA: A dataset for diverse explainable multi-hop question answering. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/1809.09600 (2018).","DOI":"10.18653\/v1\/D18-1259"},{"key":"e_1_3_3_1_58_2","unstructured":"Li Yin and Zhangyang Wang. 2025. LLM-AutoDiff: Auto-Differentiate Any LLM Workflow. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2501.16673 (2025)."},{"key":"e_1_3_3_1_59_2","unstructured":"Huaiyuan Ying Shuo Zhang Linyang Li Zhejian Zhou Yunfan Shao Zhaoye Fei Yichuan Ma Jiawei Hong Kuikun Liu Ziyi Wang et\u00a0al. 2024. Internlm-math: Open math large language models toward verifiable reasoning. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2402.06332 (2024)."},{"key":"e_1_3_3_1_60_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-1-84996-129-5"},{"key":"e_1_3_3_1_61_2","doi-asserted-by":"crossref","unstructured":"Mert Yuksekgonul Federico Bianchi Joseph Boen Sheng Liu Pan Lu Zhi Huang Carlos Guestrin and James Zou. 2025. Optimizing generative AI by backpropagating language model feedback. Nature 639 (2025) 609\u2013616.","DOI":"10.1038\/s41586-025-08661-4"},{"key":"e_1_3_3_1_62_2","doi-asserted-by":"crossref","unstructured":"Lianmin Zheng Liangsheng Yin Zhiqiang Xie Chuyue\u00a0Livia Sun Jeff Huang Cody\u00a0Hao Yu Shiyi Cao Christos Kozyrakis Ion Stoica Joseph\u00a0E Gonzalez et\u00a0al. 2024. Sglang: Efficient execution of structured language model programs. Advances in neural information processing systems 37 (2024) 62557\u201362583.","DOI":"10.52202\/079017-2000"},{"key":"e_1_3_3_1_63_2","unstructured":"Han Zhou Xingchen Wan Ruoxi Sun Hamid Palangi Shariq Iqbal Ivan Vuli\u0107 Anna Korhonen and Sercan\u00a0\u00d6 Ar\u0131k. 2025. Multi-agent design: Optimizing agents with better prompts and topologies. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2502.02533 (2025)."},{"key":"e_1_3_3_1_64_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2024.acl-long.155"},{"key":"e_1_3_3_1_65_2","volume-title":"The Eleventh International Conference on Learning Representations","author":"Zhou Yongchao","year":"2022","unstructured":"Yongchao Zhou, Andrei\u00a0Ioan Muresanu, Ziwen Han, Keiran Paster, Silviu Pitis, Harris Chan, and Jimmy Ba. 2022. Large language models are human-level prompt engineers. In The Eleventh International Conference on Learning Representations."}],"event":{"name":"CAIS '26: ACM Conference on AI and Agentic Systems","location":"San Jose CA USA","acronym":"CAIS '26"},"container-title":["Proceedings of the ACM Conference on AI and Agentic Systems"],"original-title":[],"deposited":{"date-parts":[[2026,5,22]],"date-time":"2026-05-22T03:26:49Z","timestamp":1779420409000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3786335.3813168"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,5,26]]},"references-count":64,"alternative-id":["10.1145\/3786335.3813168","10.1145\/3786335"],"URL":"https:\/\/doi.org\/10.1145\/3786335.3813168","relation":{},"subject":[],"published":{"date-parts":[[2026,5,26]]},"assertion":[{"value":"2026-05-26","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}