{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,11]],"date-time":"2026-07-11T02:27:50Z","timestamp":1783736870027,"version":"3.55.0"},"reference-count":81,"publisher":"Association for Computing Machinery (ACM)","issue":"6","content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Proc. ACM Manag. Data"],"published-print":{"date-parts":[[2025,12,4]]},"abstract":"<jats:p>Manually conducting real-world data analyses is labor-intensive and inefficient. Despite numerous attempts to automate data science workflows, none of the existing paradigms or systems fully demonstrate all three key capabilities required to support them effectively: (1) open-domain data collection, (2) structured data transformation, and (3) analytic reasoning.<\/jats:p>\n                  <jats:p>\n                    To overcome these limitations, we propose\n                    <jats:sc>Drama<\/jats:sc>\n                    , an end-to-end paradigm that answers users' analytic queries in natural language on large-scale open-domain data.\n                    <jats:sc>Drama<\/jats:sc>\n                    unifies data collection, transformation, and analysis as a single pipeline. To quantitatively evaluate system performance on tasks representative of\n                    <jats:sc>Drama<\/jats:sc>\n                    , we construct a benchmark,\n                    <jats:sc>DramaBench<\/jats:sc>\n                    , consisting of two categories of tasks: claim verification and question answering, each comprising 100 instances. These tasks are derived from real-world applications that have gained significant public attention and require the retrieval and analysis of open-domain data. We develop\n                    <jats:sc>DramaBot<\/jats:sc>\n                    , a multi-agent system designed following\n                    <jats:sc>Drama<\/jats:sc>\n                    . It comprises a data retriever that collects and transforms data by coordinating the execution of sub-agents, and a data analyzer that performs structured reasoning over the retrieved data. We evaluate\n                    <jats:sc>DramaBot<\/jats:sc>\n                    on\n                    <jats:sc>DramaBench<\/jats:sc>\n                    together with five state-of-the-art baseline agents.\n                    <jats:sc>DramaBot<\/jats:sc>\n                    achieves 86.5% task accuracy at a cost of $0.05, outperforming all baselines with up to 6.9 times the accuracy and less than 1\/6 of the cost.\n                    <jats:sc>Drama<\/jats:sc>\n                    is publicly available at https:\/\/github.com\/uiuc-kang-lab\/drama.\n                  <\/jats:p>","DOI":"10.1145\/3769781","type":"journal-article","created":{"date-parts":[[2025,12,6]],"date-time":"2025-12-06T04:32:13Z","timestamp":1764995533000},"page":"1-28","source":"Crossref","is-referenced-by-count":1,"title":["<scp>Drama<\/scp>\n                    : Unifying Data Retrieval and Analysis for Open-Domain Analytic Queries"],"prefix":"10.1145","volume":"3","author":[{"ORCID":"https:\/\/orcid.org\/0009-0001-3746-2722","authenticated-orcid":false,"given":"Chuxuan","family":"Hu","sequence":"first","affiliation":[{"name":"University of Illinois Urbana-Champaign, Champaign, IL, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0005-2293-6736","authenticated-orcid":false,"given":"Maxwell","family":"Yang","sequence":"additional","affiliation":[{"name":"University of Illinois Urbana-Champaign, Champaign, IL, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0009-3856-9247","authenticated-orcid":false,"given":"James","family":"Weiland","sequence":"additional","affiliation":[{"name":"University of Illinois Urbana-Champaign, Champaign, IL, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0000-8044-2244","authenticated-orcid":false,"given":"Yeji","family":"Lim","sequence":"additional","affiliation":[{"name":"University of Illinois Urbana-Champaign, Champaign, IL, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0002-5041-6488","authenticated-orcid":false,"given":"Suhas","family":"Palawala","sequence":"additional","affiliation":[{"name":"University of Illinois Urbana-Champaign, Champaign, IL, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-9860-9938","authenticated-orcid":false,"given":"Daniel","family":"Kang","sequence":"additional","affiliation":[{"name":"University of Illinois Urbana-Champaign, Champaign, IL, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2025,12,5]]},"reference":[{"key":"e_1_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/TVCG.2018.2865040"},{"key":"e_1_2_1_2_1","unstructured":"Asim Biswal Liana Patel Siddarth Jha Amog Kamsetty Shu Liu Joseph E. Gonzalez Carlos Guestrin and Matei Zaharia. 2024. Text2SQL is Not Enough: Unifying AI and Databases with TAG. arXiv:2408.14717 [cs.DB] https:\/\/arxiv.org\/abs\/2408.14717"},{"key":"e_1_2_1_3_1","unstructured":"Anjanava Biswas and Wrick Talukdar. 2024. Robustness of Structured Data Extraction from In-plane Rotated Documents using Multi-Modal Large Language Models (LLM). arXiv:2406.10295 [cs.CL] https:\/\/arxiv.org\/abs\/2406.10295"},{"key":"e_1_2_1_4_1","unstructured":"Tom B. Brown Benjamin Mann Nick Ryder Melanie Subbiah Jared Kaplan Prafulla Dhariwal Arvind Neelakantan Pranav Shyam Girish Sastry Amanda Askell Sandhini Agarwal Ariel Herbert-Voss Gretchen Krueger Tom Henighan Rewon Child Aditya Ramesh Daniel M. Ziegler Jeffrey Wu Clemens Winter Christopher Hesse Mark Chen Eric Sigler Mateusz Litwin Scott Gray Benjamin Chess Jack Clark Christopher Berner Sam McCandlish Alec Radford Ilya Sutskever and Dario Amodei. 2020. Language Models are Few-Shot Learners. arXiv:2005.14165 [cs.CL] https:\/\/arxiv.org\/abs\/2005.14165"},{"key":"e_1_2_1_5_1","unstructured":"Ruisheng Cao Fangyu Lei Haoyuan Wu Jixuan Chen Yeqiao Fu Hongcheng Gao Xinzhuang Xiong Hanchong Zhang Yuchen Mao Wenjing Hu Tianbao Xie Hongshen Xu Danyang Zhang Sida Wang Ruoxi Sun Pengcheng Yin Caiming Xiong Ansong Ni Qian Liu Victor Zhong Lu Chen Kai Yu and Tao Yu. 2024. Spider2-V: How Far Are Multimodal Agents From Automating Data Science and Engineering Workflows? arXiv:2407.10956 [cs.AI] https:\/\/arxiv.org\/abs\/2407.10956"},{"key":"e_1_2_1_6_1","volume-title":"Vital Statistics Yearly Summary","author":"Office Ireland Central Statistics","year":"2023","unstructured":"Central Statistics Office Ireland. 2024. Vital Statistics Yearly Summary 2023. https:\/\/www.cso.ie\/en\/releasesandpublications\/ep\/p-vsys\/vitalstatisticsyearlysummary2023\/"},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1145\/3583780.3614821"},{"key":"e_1_2_1_8_1","unstructured":"Mark Chen Jerry Tworek Heewoo Jun Qiming Yuan Henrique Ponde de Oliveira Pinto Jared Kaplan Harri Edwards Yuri Burda Nicholas Joseph Greg Brockman Alex Ray Raul Puri Gretchen Krueger Michael Petrov Heidy Khlaaf Girish Sastry Pamela Mishkin Brooke Chan Scott Gray Nick Ryder Mikhail Pavlov Alethea Power Lukasz Kaiser Mohammad Bavarian Clemens Winter Philippe Tillet Felipe Petroski Such Dave Cummings Matthias Plappert Fotios Chantzis Elizabeth Barnes Ariel Herbert-Voss William Hebgen Guss Alex Nichol Alex Paino Nikolas Tezak Jie Tang Igor Babuschkin Suchir Balaji Shantanu Jain William Saunders Christopher Hesse Andrew N. Carr Jan Leike Josh Achiam Vedant Misra Evan Morikawa Alec Radford Matthew Knight Miles Brundage Mira Murati Katie Mayer Peter Welinder Bob McGrew Dario Amodei Sam McCandlish Ilya Sutskever and Wojciech Zaremba. 2021. Evaluating Large Language Models Trained on Code. arXiv:2107.03374 [cs.LG] https:\/\/arxiv.org\/abs\/2107.03374"},{"key":"e_1_2_1_9_1","unstructured":"Community Notes. 2025. Download Data - Community Notes Guide. https:\/\/communitynotes.x.com\/guide\/en\/under-the-hood\/download-data. https:\/\/communitynotes.x.com\/guide\/en\/under-the-hood\/download-data"},{"key":"e_1_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2020.findings-emnlp.27"},{"key":"e_1_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1145\/988672.988687"},{"key":"e_1_2_1_12_1","volume-title":"Proceedings of the Twenty-Second International Joint Conference on Artificial Intelligence -","volume":"10","author":"Etzioni Oren","year":"2011","unstructured":"Oren Etzioni, Anthony Fader, Janara Christensen, Stephen Soderland, and Mausam Mausam. 2011. Open information extraction: the second generation. In Proceedings of the Twenty-Second International Joint Conference on Artificial Intelligence - Volume Volume One (Barcelona, Catalonia, Spain) (IJCAI'11). AAAI Press, 3-10."},{"key":"e_1_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2022.naacl-main.187"},{"key":"e_1_2_1_14_1","unstructured":"Federation for American Immigration Reform (FAIR). 2023. 2023 Illegal Alien Population Estimate. https:\/\/www.fairus.org\/sites\/default\/files\/2023-06\/2023%20Illegal%20Alien%20Population%20Estimate_2.pdf."},{"key":"e_1_2_1_15_1","unstructured":"Federation for American Immigration Reform (FAIR). 2025. Federation for American Immigration Reform. https:\/\/www.fairus.org."},{"key":"e_1_2_1_16_1","doi-asserted-by":"crossref","unstructured":"Zhangyin Feng Daya Guo Duyu Tang Nan Duan Xiaocheng Feng Ming Gong Linjun Shou Bing Qin Ting Liu Daxin Jiang and Ming Zhou. 2020. CodeBERT: A Pre-Trained Model for Programming and Natural Languages. arXiv:2002.08155 [cs.CL] https:\/\/arxiv.org\/abs\/2002.08155","DOI":"10.18653\/v1\/2020.findings-emnlp.139"},{"key":"e_1_2_1_17_1","volume-title":"Text-to-SQL Empowered by Large Language Models: A Benchmark Evaluation. CoRR","author":"Gao Dawei","year":"2023","unstructured":"Dawei Gao, Haibin Wang, Yaliang Li, Xiuyu Sun, Yichen Qian, Bolin Ding, and Jingren Zhou. 2023b. Text-to-SQL Empowered by Large Language Models: A Benchmark Evaluation. CoRR, Vol. abs\/2308.15363 (2023)."},{"key":"e_1_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2023.acl-long.910"},{"key":"e_1_2_1_19_1","unstructured":"Google. 2024. Gemini Deep Research. https:\/\/gemini.google\/overview\/deep-research\/."},{"key":"e_1_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2021.naacl-main.114"},{"key":"e_1_2_1_21_1","unstructured":"Significant Gravitas. 2023. Auto-GPT: An Autonomous GPT-4 Experiment. https:\/\/github.com\/Significant-Gravitas\/AutoGPT GitHub repository."},{"key":"e_1_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-24758-3_10"},{"key":"e_1_2_1_23_1","unstructured":"Hongliang He Wenlin Yao Kaixin Ma Wenhao Yu Yong Dai Hongming Zhang Zhenzhong Lan and Dong Yu. 2024. WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models. arXiv:2401.13919 [cs.CL] https:\/\/arxiv.org\/abs\/2401.13919"},{"key":"e_1_2_1_24_1","volume-title":"Data Interpreter: An LLM Agent For Data Science. arXiv:2402.18679 [cs.AI] https:\/\/arxiv.org\/abs\/2402.18679","author":"Hong Sirui","year":"2024","unstructured":"Sirui Hong, Yizhang Lin, Bang Liu, Bangbang Liu, Binhao Wu, Ceyao Zhang, Chenxing Wei, Danyang Li, Jiaqi Chen, Jiayi Zhang, Jinlin Wang, Li Zhang, Lingyao Zhang, Min Yang, Mingchen Zhuge, Taicheng Guo, Tuo Zhou, Wei Tao, Xiangru Tang, Xiangtao Lu, Xiawu Zheng, Xinbing Liang, Yaying Fei, Yuheng Cheng, Zhibin Gou, Zongze Xu, and Chenglin Wu. 2024. Data Interpreter: An LLM Agent For Data Science. arXiv:2402.18679 [cs.AI] https:\/\/arxiv.org\/abs\/2402.18679"},{"key":"e_1_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.14778\/3705829.3705843"},{"key":"e_1_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2025.findings-acl.1210"},{"key":"e_1_2_1_27_1","unstructured":"Indian Tech & Infra. 2024. Uttar Pradesh is now the second-largest economy in the country competing to become India's top economy: CM Yogi Adityanath. https:\/\/x.com\/IndianTechGuide\/status\/1850054756921925792."},{"key":"e_1_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2023.emnlp-main.470"},{"key":"e_1_2_1_29_1","unstructured":"Serafina Kamp Morteza Fayazi Zineb Benameur-El Shuyan Yu and Ronald Dreslinski. 2023. Open Information Extraction: A Review of Baseline Techniques Approaches and Applications. arXiv:2310.11644 [cs.IR] https:\/\/arxiv.org\/abs\/2310.11644"},{"key":"e_1_2_1_30_1","volume-title":"Akari Asai, Xinyan Yu, Dragomir Radev, Noah A. Smith, Yejin Choi, and Kentaro Inui.","author":"Kasai Jungo","year":"2024","unstructured":"Jungo Kasai, Keisuke Sakaguchi, Yoichi Takahashi, Ronan Le Bras, Akari Asai, Xinyan Yu, Dragomir Radev, Noah A. Smith, Yejin Choi, and Kentaro Inui. 2024. RealTime QA: What's the Answer Right Now? arXiv:2207.13332 [cs.CL] https:\/\/arxiv.org\/abs\/2207.13332"},{"key":"e_1_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2020.emnlp-main.750"},{"key":"e_1_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1162\/tacl_a_00453"},{"key":"e_1_2_1_33_1","unstructured":"Fangyu Lei Jixuan Chen Yuxiao Ye Ruisheng Cao Dongchan Shin Hongjin Su Zhaoqing Suo Hongcheng Gao Wenjing Hu Pengcheng Yin et al. 2024. Spider 2.0: Evaluating language models on real-world enterprise text-to-sql workflows. arXiv preprint arXiv:2411.07763 (2024)."},{"key":"e_1_2_1_34_1","unstructured":"Fangyu Lei Jixuan Chen Yuxiao Ye Ruisheng Cao Dongchan Shin Hongjin Su Zhaoqing Suo Hongcheng Gao Wenjing Hu Pengcheng Yin Victor Zhong Caiming Xiong Ruoxi Sun Qian Liu Sida Wang and Tao Yu. 2025. Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows. arXiv:2411.07763 [cs.CL] https:\/\/arxiv.org\/abs\/2411.07763"},{"key":"e_1_2_1_35_1","volume-title":"Tim Rockt\u00e4schel, Sebastian Riedel, and Douwe Kiela.","author":"Lewis Patrick","year":"2021","unstructured":"Patrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich K\u00fcttler, Mike Lewis, Wen tau Yih, Tim Rockt\u00e4schel, Sebastian Riedel, and Douwe Kiela. 2021. Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks. arXiv:2005.11401 [cs.CL] https:\/\/arxiv.org\/abs\/2005.11401"},{"key":"e_1_2_1_36_1","volume-title":"Advances in Neural Information Processing Systems","volume":"36","author":"Li Jinyang","year":"2024","unstructured":"Jinyang Li, Binyuan Hui, Ge Qu, Jiaxi Yang, Binhua Li, Bowen Li, Bailin Wang, Bowen Qin, Ruiying Geng, Nan Huo, et al., 2024. Can llm already serve as a database interface? a big bench for large-scale database grounded text-to-sqls. Advances in Neural Information Processing Systems, Vol. 36 (2024)."},{"key":"e_1_2_1_37_1","volume-title":"Zui Chen, Michael Franklin, Tim Kraska, Samuel Madden, and Gerardo Vitagliano.","author":"Liu Chunwei","year":"2024","unstructured":"Chunwei Liu, Matthew Russo, Michael Cafarella, Lei Cao, Peter Baille Chen, Zui Chen, Michael Franklin, Tim Kraska, Samuel Madden, and Gerardo Vitagliano. 2024b. A Declarative System for Optimizing AI Workloads. arXiv:2405.14696 [cs.CL] https:\/\/arxiv.org\/abs\/2405.14696"},{"key":"e_1_2_1_38_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2023.findings-emnlp.467"},{"key":"e_1_2_1_39_1","doi-asserted-by":"publisher","DOI":"10.1162\/tacl_a_00638"},{"key":"e_1_2_1_40_1","volume-title":"Phil Blunsom, and Angeliki Lazaridou.","author":"Liska Adam","year":"2022","unstructured":"Adam Liska, Elena Gribovskaya, Tayfun Terzi, Eren Sezener, Devang Agrawal, Cyprien de Masson d'Autume, Tim Scholtes, Manzil Zaheer, Susannah Young, Ellen Gilsenan-McMahon, Sophia Austin, Phil Blunsom, and Angeliki Lazaridou. 2022. StreamingQA: A Benchmark for Adaptation to New Knowledge over Time in Question Answering Models. arXiv:2205.11388 [cs.CL] https:\/\/arxiv.org\/abs\/2205.11388"},{"key":"e_1_2_1_41_1","doi-asserted-by":"crossref","unstructured":"Chaitanya Malaviya Subin Lee Sihao Chen Elizabeth Sieber Mark Yatskar and Dan Roth. 2024. ExpertQA: Expert-Curated Questions and Attributed Answers. arXiv:2309.07852 [cs.CL] https:\/\/arxiv.org\/abs\/2309.07852","DOI":"10.18653\/v1\/2024.naacl-long.167"},{"key":"e_1_2_1_42_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2023.emnlp-main.741"},{"key":"e_1_2_1_43_1","unstructured":"National Park Service. 2023. Visitor Spending Effects: Economic Contributions of National Park Visitor Spending. https:\/\/www.nps.gov\/nature\/customcf\/NPS_Data_Visualization\/docs\/NPS_2023_Visitor_Spending_Effects.pdf"},{"key":"e_1_2_1_44_1","unstructured":"National Park Service. 2025. Nature - U.S. National Park Service. https:\/\/www.nps.gov\/nature\/. https:\/\/www.nps.gov\/nature\/"},{"key":"e_1_2_1_45_1","unstructured":"OpenAI. 2024a. GPT-4o mini. https:\/\/platform.openai.com\/docs\/models\/gpt-4o-mini."},{"key":"e_1_2_1_46_1","unstructured":"OpenAI. 2024b. GPT-4o Overview. https:\/\/platform.openai.com\/docs\/models\/gpt-4o. Accessed: 2025-04-11."},{"key":"e_1_2_1_47_1","unstructured":"OpenAI. 2024. Introducing GPT-4o. https:\/\/openai.com\/index\/hello-gpt-4o\/. Accessed: 2025-08-03."},{"key":"e_1_2_1_48_1","unstructured":"OpenAI. 2024. O3 Mini Model Documentation. https:\/\/platform.openai.com\/docs\/models\/o3-mini."},{"key":"e_1_2_1_49_1","unstructured":"OpenAI. 2024. OpenAI Agents Python: Research Bot Example. https:\/\/github.com\/openai\/openai-agents-python\/tree\/main\/examples\/research_bot. Accessed: 2025-04-08."},{"key":"e_1_2_1_50_1","unstructured":"OpenAI. 2024. OpenAI Embeddings Guide. https:\/\/platform.openai.com\/docs\/guides\/embeddings."},{"key":"e_1_2_1_51_1","unstructured":"OpenAI. 2024. Tools and Web Search | OpenAI Platform. https:\/\/platform.openai.com\/docs\/guides\/tools-web-search?api-mode=chat Accessed: 2025-04-08."},{"key":"e_1_2_1_52_1","volume-title":"Semantic Operators: A Declarative Model for Rich, AI-based Data Processing. arXiv:2407.11418 [cs.DB] https:\/\/arxiv.org\/abs\/2407.11418","author":"Patel Liana","year":"2025","unstructured":"Liana Patel, Siddharth Jha, Melissa Pan, Harshit Gupta, Parth Asawa, Carlos Guestrin, and Matei Zaharia. 2025. Semantic Operators: A Declarative Model for Rich, AI-based Data Processing. arXiv:2407.11418 [cs.DB] https:\/\/arxiv.org\/abs\/2407.11418"},{"key":"e_1_2_1_53_1","doi-asserted-by":"publisher","DOI":"10.1098\/rspl.1895.0041"},{"key":"e_1_2_1_54_1","volume-title":"ast - Abstract Syntax Trees","author":"Foundation Python Software","unstructured":"Python Software Foundation. 2024. ast - Abstract Syntax Trees. Python Software Foundation. https:\/\/docs.python.org\/3\/library\/ast.html"},{"key":"e_1_2_1_55_1","volume-title":"Selenium: Web Browser Automation. https:\/\/www.selenium.dev\/. https:\/\/www.selenium.dev\/ Accessed: 2025-08-04.","author":"Project Selenium","year":"2025","unstructured":"Selenium Project. 2025. Selenium: Web Browser Automation. https:\/\/www.selenium.dev\/. https:\/\/www.selenium.dev\/ Accessed: 2025-08-04."},{"key":"e_1_2_1_56_1","doi-asserted-by":"crossref","unstructured":"Shreya Shankar Tristan Chambers Tarak Shah Aditya G. Parameswaran and Eugene Wu. 2025. DocETL: Agentic Query Rewriting and Evaluation for Complex Document Processing. arXiv:2410.12189 [cs.DB] https:\/\/arxiv.org\/abs\/2410.12189","DOI":"10.14778\/3746405.3746426"},{"key":"e_1_2_1_57_1","volume-title":"India: PM Air Pollution Exposure. https:\/\/www.statista.com\/statistics\/935636\/india-pm-air-pollution-exposure\/ Accessed: 2025-04-14.","year":"2024","unstructured":"Statista. 2024. India: PM Air Pollution Exposure. https:\/\/www.statista.com\/statistics\/935636\/india-pm-air-pollution-exposure\/ Accessed: 2025-04-14."},{"key":"e_1_2_1_58_1","unstructured":"Statista. 2025. The U.S. Minimum Wage by State. https:\/\/www.statista.com\/statistics\/238997\/minimum-wage-by-us-state\/."},{"key":"e_1_2_1_59_1","unstructured":"Statistics Times. 2025. List of Indian States and Union Territories by GDP. https:\/\/statisticstimes.com\/economy\/india\/indian-states-gdp.php."},{"key":"e_1_2_1_60_1","doi-asserted-by":"publisher","DOI":"10.1145\/3360646"},{"key":"e_1_2_1_61_1","unstructured":"Bashir Tahir. 2024. Open Deep Research. https:\/\/github.com\/btahir\/open-deep-research. Accessed: 2025-04-08."},{"key":"e_1_2_1_62_1","doi-asserted-by":"crossref","unstructured":"Liyan Tang Philippe Laban and Greg Durrett. 2024a. MiniCheck: Efficient Fact-Checking of LLMs on Grounding Documents. arXiv:2404.10774 [cs.CL] https:\/\/arxiv.org\/abs\/2404.10774","DOI":"10.18653\/v1\/2024.emnlp-main.499"},{"key":"e_1_2_1_63_1","volume-title":"Jon Burnsky, Jake W. Vincent, Yu'an Yang, Siffi Singh, Song Feng, Hwanjun Song, Hang Su, Lijia Sun, Yi Zhang, Saab Mansour, and Kathleen McKeown.","author":"Tang Liyan","year":"2024","unstructured":"Liyan Tang, Igor Shalyminov, Amy Wing mei Wong, Jon Burnsky, Jake W. Vincent, Yu'an Yang, Siffi Singh, Song Feng, Hwanjun Song, Hang Su, Lijia Sun, Yi Zhang, Saab Mansour, and Kathleen McKeown. 2024b. TofuEval: Evaluating Hallucinations of LLMs on Topic-Focused Dialogue Summarization. arXiv:2402.13249 [cs.CL] https:\/\/arxiv.org\/abs\/2402.13249"},{"key":"e_1_2_1_64_1","unstructured":"Maria Ramirez Uribe. 2024. There aren't 20 million to 30 million immigrants in the U.S. illegally as Sen. Marco Rubio claimed. https:\/\/www.politifact.com\/factchecks\/2024\/jun\/11\/marco-rubio\/there-arent-20-million-to-30-million-immigrants-in\/."},{"key":"e_1_2_1_65_1","unstructured":"U.S. Department of Housing and Urban Development. 2024. 2024 AHAR: Part 1 - PIT Estimates of Homelessness in the U.S. https:\/\/www.huduser.gov\/portal\/datasets\/ahar\/2024-ahar-part-1-pit-estimates-of-homelessness-in-the-us.html"},{"key":"e_1_2_1_66_1","unstructured":"USAFacts. 2024a. How do national parks affect the economy? https:\/\/usafacts.org\/articles\/how-do-national-parks-affect-the-economy\/"},{"key":"e_1_2_1_67_1","unstructured":"USAFacts. 2024b. How much damage do wildfires do in the US? https:\/\/usafacts.org\/articles\/how-much-damage-do-wildfires-do-in-the-us\/ Accessed: 2025-04-14."},{"key":"e_1_2_1_68_1","unstructured":"USAFacts. 2024c. Which states have the highest and lowest rates of homelessness? https:\/\/usafacts.org\/articles\/which-states-have-the-highest-and-lowest-rates-of-homelessness\/"},{"key":"e_1_2_1_69_1","unstructured":"USAFacts. 2025a. Minimum wage in America: How many people are earning $7.25 an hour? https:\/\/usafacts.org\/articles\/minimum-wage-america-how-many-people-are-earning-725-hour\/."},{"key":"e_1_2_1_70_1","unstructured":"USAFacts. 2025b. USAFacts: Nonpartisan data. Driven by facts. https:\/\/usafacts.org\/. https:\/\/usafacts.org\/"},{"key":"e_1_2_1_71_1","unstructured":"USAFacts. 2025c. USAFacts: Nonpartisan Government Data. https:\/\/usafacts.org\/. https:\/\/usafacts.org\/ Accessed: 2025-08-04."},{"key":"e_1_2_1_72_1","unstructured":"Xintao Wang Qianwen Yang Yongting Qiu Jiaqing Liang Qianyu He Zhouhong Gu Yanghua Xiao and Wei Wang. 2023. KnowledGPT: Enhancing Large Language Models with Retrieval and Storage Access on Knowledge Bases. arXiv:2308.11761 [cs.CL] https:\/\/arxiv.org\/abs\/2308.11761"},{"key":"e_1_2_1_73_1","volume-title":"Zain Muhammad Mujahid, Arnav Arora, Aleksandr Rubashevskii, Jiahui Geng, Osama Mohammed Afzal, Liangming Pan, Nadav Borenstein, Aditya Pillai, Isabelle Augenstein, Iryna Gurevych, and Preslav Nakov.","author":"Wang Yuxia","year":"2024","unstructured":"Yuxia Wang, Revanth Gangi Reddy, Zain Muhammad Mujahid, Arnav Arora, Aleksandr Rubashevskii, Jiahui Geng, Osama Mohammed Afzal, Liangming Pan, Nadav Borenstein, Aditya Pillai, Isabelle Augenstein, Iryna Gurevych, and Preslav Nakov. 2024. Factcheck-Bench: Fine-Grained Evaluation Benchmark for Automatic Fact-checkers. arXiv:2311.09000 [cs.CL] https:\/\/arxiv.org\/abs\/2311.09000"},{"key":"e_1_2_1_74_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2023.findings-emnlp.167"},{"key":"e_1_2_1_75_1","volume-title":"Education Rankings by Country","author":"Review World Population","year":"2024","unstructured":"World Population Review. 2024. Education Rankings by Country 2024. https:\/\/worldpopulationreview.com\/country-rankings\/education-rankings-by-country Accessed: 2025-04-14."},{"key":"e_1_2_1_76_1","unstructured":"X (formerly Twitter). 2024a. Community Note clarifying claim on birth rate in Ireland. https:\/\/x.com\/i\/birdwatch\/n\/1821558069928620304 Accessed: 2025-04-14."},{"key":"e_1_2_1_77_1","unstructured":"X (formerly Twitter). 2024b. Post on X referencing change in number of births in Ireland (2022-2023). https:\/\/x.com\/i\/web\/status\/1821510949771059296"},{"key":"e_1_2_1_78_1","unstructured":"X (formerly Twitter). 2024c. Post referencing global education rankings. https:\/\/x.com\/i\/web\/status\/1769443227902341380 Accessed: 2025-04-14."},{"key":"e_1_2_1_79_1","unstructured":"X (formerly Twitter). 2024d. Post referencing India's cigarette consumption statistic. https:\/\/x.com\/i\/web\/status\/1858383505501098277 Accessed: 2025-04-14."},{"key":"e_1_2_1_80_1","unstructured":"Hui Yang Sifu Yue and Yunzhong He. 2023. Auto-GPT for Online Decision Making: Benchmarks and Additional Opinions. arXiv:2306.02224 [cs.AI] https:\/\/arxiv.org\/abs\/2306.02224"},{"key":"e_1_2_1_81_1","unstructured":"Wenqi Zhang Yongliang Shen Weiming Lu and Yueting Zhuang. 2024. Data-Copilot: Bridging Billions of Data and Humans with Autonomous Workflow. arXiv:2306.07209 [cs.CL] https:\/\/arxiv.org\/abs\/2306.07209"}],"container-title":["Proceedings of the ACM on Management of Data"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3769781","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,6,13]],"date-time":"2026-06-13T04:51:04Z","timestamp":1781326264000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3769781"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,12,4]]},"references-count":81,"journal-issue":{"issue":"6","published-print":{"date-parts":[[2025,12,4]]}},"alternative-id":["10.1145\/3769781"],"URL":"https:\/\/doi.org\/10.1145\/3769781","relation":{},"ISSN":["2836-6573"],"issn-type":[{"value":"2836-6573","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,12,4]]}}}