{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,11,19]],"date-time":"2025-11-19T19:28:49Z","timestamp":1763580529780,"version":"3.44.0"},"reference-count":63,"publisher":"Association for Computing Machinery (ACM)","issue":"CSCW2","license":[{"start":{"date-parts":[[2024,11,7]],"date-time":"2024-11-07T00:00:00Z","timestamp":1730937600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["Proc. ACM Hum.-Comput. Interact."],"published-print":{"date-parts":[[2024,11,7]]},"abstract":"<jats:p>Stack Overflow (SO) is a widely recognized online question-and-answer platform for programming, which has also fostered a substantial community dedicated to machine learning (ML), providing a space for both novices and experts to exchange ideas and find solutions to ML-related problems. However, as a relative minority of this online programming platform, research has demonstrated lower engagement in the ML community, but it remains largely unexplored to understand what hinders the engagement and contribution from ML users' perspectives. This paper presents an empirical study based on 22 hours of semi-structured interviews and 131 survey responses with users on SO and reveals the key factors that may lead to the lower response rate and extended waiting time for ML questions on SO, which includes the unique quality requirement for posting ML questions, the discrepancy between time invested and benefits gained, the dispersed nature of the ML community across various platforms and the desired improvement for SO. Moreover, the qualitative study reveals a declining friendliness in SO's culture over time; the subsequent quantitative study corroborates that newcomers frequently encounter stress when posting and answering ML questions, even though this stress diminishes with increased experience. Additionally, we also explored the potential influence of generative AI tools (e.g., ChatGPT) on online question-and-answer platforms, specifically focusing on ML Q&amp;A. We hope the results of this study can pave the way for enhancing the experience of ML users on online platforms, ultimately facilitating improved knowledge exchange and collaboration within the ML domain.<\/jats:p>","DOI":"10.1145\/3686990","type":"journal-article","created":{"date-parts":[[2024,11,8]],"date-time":"2024-11-08T15:52:40Z","timestamp":1731081160000},"page":"1-35","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":1,"title":["\"Math is a pain!\": Understanding challenges and needs of the Machine Learning community on Stack Overflow"],"prefix":"10.1145","volume":"8","author":[{"ORCID":"https:\/\/orcid.org\/0009-0009-2151-2922","authenticated-orcid":false,"given":"Zihan","family":"Fang","sequence":"first","affiliation":[{"name":"Vanderbilt University, Nashville, TN, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-2730-5077","authenticated-orcid":false,"given":"Yu","family":"Huang","sequence":"additional","affiliation":[{"name":"Vanderbilt University, Nashville, TN, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2024,11,8]]},"reference":[{"key":"e_1_2_2_1_1","doi-asserted-by":"publisher","DOI":"10.1145\/2901739.2901770"},{"key":"e_1_2_2_2_1","volume-title":"Database","volume":"2020","author":"Ahmed Zeeshan","year":"2020","unstructured":"Zeeshan Ahmed, Khalid Mohamed, Saman Zeeshan, and XinQi Dong. 2020. Artificial intelligence with multi-functional machine learning platform development for better healthcare and precision medicine. Database, Vol. 2020 (2020), baaa010."},{"key":"e_1_2_2_3_1","doi-asserted-by":"publisher","DOI":"10.3390\/educsci11090552"},{"volume-title":"Why is developing machine learning applications challenging? a study on stack overflow posts. In 2019 acm\/ieee international symposium on empirical software engineering and measurement (esem)","author":"Alshangiti Moayad","key":"e_1_2_2_4_1","unstructured":"Moayad Alshangiti, Hitesh Sapkota, Pradeep K Murukannaiah, Xumin Liu, and Qi Yu. 2019. Why is developing machine learning applications challenging? a study on stack overflow posts. In 2019 acm\/ieee international symposium on empirical software engineering and measurement (esem). IEEE, 1--11."},{"key":"e_1_2_2_5_1","volume-title":"Workshop report on basic research needs for scientific machine learning: Core technologies for artificial intelligence. Technical Report. USDOE Office of Science (SC), Washington, DC (United States).","author":"Baker Nathan","year":"2019","unstructured":"Nathan Baker, Frank Alexander, Timo Bremer, Aric Hagberg, Yannis Kevrekidis, Habib Najm, Manish Parashar, Abani Patra, James Sethian, Stefan Wild, et al. 2019. Workshop report on basic research needs for scientific machine learning: Core technologies for artificial intelligence. Technical Report. USDOE Office of Science (SC), Washington, DC (United States)."},{"key":"e_1_2_2_6_1","volume-title":"What are developers talking about? an analysis of topics and trends in stack overflow. Empirical software engineering","author":"Barua Anton","year":"2014","unstructured":"Anton Barua, Stephen W Thomas, and Ahmed E Hassan. 2014. What are developers talking about? an analysis of topics and trends in stack overflow. Empirical software engineering, Vol. 19 (2014), 619--654."},{"volume-title":"Qualitative HCI research: Going behind the scenes","author":"Blandford Ann","key":"e_1_2_2_7_1","unstructured":"Ann Blandford, Dominic Furniss, and Stephann Makri. 2016. Qualitative HCI research: Going behind the scenes. Morgan & Claypool Publishers."},{"volume-title":"Thematic analysis","author":"Braun Virginia","key":"e_1_2_2_8_1","unstructured":"Virginia Braun and Victoria Clarke. 2012. Thematic analysis.American Psychological Association."},{"key":"e_1_2_2_9_1","doi-asserted-by":"publisher","DOI":"10.1145\/3134667"},{"key":"e_1_2_2_10_1","doi-asserted-by":"crossref","unstructured":"Yinong Chen. 2020. IoT cloud big data and AI in interdisciplinary domains. 102070 pages.","DOI":"10.1016\/j.simpat.2020.102070"},{"key":"e_1_2_2_11_1","unstructured":"Henry Clay. [n. d.]. The Paired t-Test and the Wilcoxon Matched-Pairs Test: Comparing the Means of Two Related Groups. for Nursing and Allied Health ( [n. d.]) 139."},{"key":"e_1_2_2_12_1","doi-asserted-by":"publisher","DOI":"10.1109\/TCSS.2022.3151130"},{"key":"e_1_2_2_13_1","doi-asserted-by":"publisher","DOI":"10.1080\/01431160500275762"},{"key":"e_1_2_2_14_1","doi-asserted-by":"publisher","DOI":"10.1145\/2556420.2556477"},{"key":"e_1_2_2_15_1","doi-asserted-by":"publisher","DOI":"10.1145\/2950290.2950331"},{"key":"e_1_2_2_16_1","doi-asserted-by":"publisher","DOI":"10.1177\/1098214011426594"},{"key":"e_1_2_2_17_1","doi-asserted-by":"publisher","DOI":"10.1037\/a0024338"},{"volume-title":"Practical machine learning","author":"Gollapudi Sunila","key":"e_1_2_2_18_1","unstructured":"Sunila Gollapudi. 2016. Practical machine learning. Packt Publishing Ltd."},{"key":"e_1_2_2_19_1","doi-asserted-by":"publisher","DOI":"10.1109\/SCAM52516.2021.00016"},{"key":"e_1_2_2_20_1","doi-asserted-by":"publisher","DOI":"10.1109\/SCAM52516.2021.00016"},{"key":"e_1_2_2_21_1","doi-asserted-by":"publisher","DOI":"10.1080\/20476965.2023.2237745"},{"key":"e_1_2_2_22_1","volume-title":"StackOverflow vs Kaggle: A Study of Developer Discussions About Data Science. arxiv","author":"Hin David","year":"2006","unstructured":"David Hin. 2020. StackOverflow vs Kaggle: A Study of Developer Discussions About Data Science. arxiv: 2006.08334 [cs.CL]"},{"key":"e_1_2_2_23_1","doi-asserted-by":"publisher","DOI":"10.1109\/METRICS.2005.24"},{"key":"e_1_2_2_24_1","volume-title":"Rangeet Pan, and Hridesh Rajan.","author":"Islam Md Johirul","year":"2019","unstructured":"Md Johirul Islam, Hoan Anh Nguyen, Rangeet Pan, and Hridesh Rajan. 2019. What do developers ask about ML libraries? A large-scale study using stack overflow. arXiv preprint arXiv:1906.11940 (2019)."},{"key":"e_1_2_2_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/TKDE.2018.2861006"},{"key":"e_1_2_2_26_1","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1155\/2021\/4535567","article-title":"Opportunities of artificial intelligence and machine learning in the food industry","volume":"2021","author":"Kumar Indrajeet","year":"2021","unstructured":"Indrajeet Kumar, Jyoti Rawat, Noor Mohd, and Shahnawaz Husain. 2021. Opportunities of artificial intelligence and machine learning in the food industry. Journal of Food Quality, Vol. 2021 (2021), 1--10.","journal-title":"Journal of Food Quality"},{"volume-title":"Reorienting machine learning education towards tinkerers and ML-engaged citizens. Ph.,D. Dissertation","author":"Lao Natalie","key":"e_1_2_2_27_1","unstructured":"Natalie Lao. 2020. Reorienting machine learning education towards tinkerers and ML-engaged citizens. Ph.,D. Dissertation. Massachusetts Institute of Technology Cambridge, MA, USA."},{"key":"e_1_2_2_28_1","unstructured":"Jun Lin. 2018. Predicting the Best Answers for Questions on Stack Overflow."},{"key":"e_1_2_2_29_1","doi-asserted-by":"publisher","DOI":"10.1109\/SANER53432.2022.00075"},{"key":"e_1_2_2_30_1","volume-title":"Proceedings of the International Conference.","author":"Louw Stephen","year":"2021","unstructured":"Stephen Louw. 2021. Automated transcription software in qualitative research. In Proceedings of the International Conference."},{"key":"e_1_2_2_31_1","unstructured":"Niklas Luhmann. 1995. Social systems. stanford university Press."},{"key":"e_1_2_2_32_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.zemedi.2018.11.002"},{"key":"e_1_2_2_33_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2022.3209825"},{"key":"e_1_2_2_34_1","doi-asserted-by":"publisher","DOI":"10.1111\/bcp.14801"},{"key":"e_1_2_2_35_1","volume-title":"Exploring Research Interest in Stack Overflow -- A Systematic Mapping Study and Quality Evaluation. arxiv","author":"Meldrum Sarah","year":"2010","unstructured":"Sarah Meldrum, Sherlock A. Licorish, and Bastin Tony Roy Savarimuthu. 2020. Exploring Research Interest in Stack Overflow -- A Systematic Mapping Study and Quality Evaluation. arxiv: 2010.12282 [cs.SE]"},{"key":"e_1_2_2_36_1","doi-asserted-by":"publisher","DOI":"10.1146\/annurev-matsci-070218-010015"},{"key":"e_1_2_2_37_1","doi-asserted-by":"publisher","DOI":"10.1371\/journal.pone.0253010"},{"key":"e_1_2_2_38_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICSM.2012.6405249"},{"key":"e_1_2_2_39_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10462-018-09679-z"},{"key":"e_1_2_2_40_1","doi-asserted-by":"publisher","DOI":"10.1145\/3533378"},{"key":"e_1_2_2_41_1","first-page":"35","article-title":"Mixed method designs: A review of strategies for blending quantitative and qualitative methodologies","volume":"20","author":"Pole Kathryn","year":"2007","unstructured":"Kathryn Pole. 2007. Mixed method designs: A review of strategies for blending quantitative and qualitative methodologies. Mid-Western Educational Researcher, Vol. 20, 4 (2007), 35--38.","journal-title":"Mid-Western Educational Researcher"},{"key":"e_1_2_2_42_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.chb.2018.06.004"},{"key":"e_1_2_2_43_1","doi-asserted-by":"publisher","unstructured":"Zishan Qin Yimin Wu Jiayan Pei Jinwei Lu Shizhao Huang and Liqun Liu. 2022. Related Questions Retrieval Model in Stack Overflow based on Semantic Matching. In 2022 IEEE 46th Annual Computers Software and Applications Conference (COMPSAC). 321--326. https:\/\/doi.org\/10.1109\/COMPSAC54236.2022.00052","DOI":"10.1109\/COMPSAC54236.2022.00052"},{"key":"e_1_2_2_44_1","doi-asserted-by":"publisher","DOI":"10.3390\/math9222970"},{"key":"e_1_2_2_45_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.matpr.2021.11.544"},{"key":"e_1_2_2_46_1","volume-title":"Enhancing Mathematical Capabilities through ChatGPT and Similar Generative Artificial Intelligence: Roles and Challenges in Solving Mathematical Problems. Available at SSRN 4603237","author":"Rane Nitin","year":"2023","unstructured":"Nitin Rane. 2023. Enhancing Mathematical Capabilities through ChatGPT and Similar Generative Artificial Intelligence: Roles and Challenges in Solving Mathematical Problems. Available at SSRN 4603237 (2023)."},{"key":"e_1_2_2_47_1","doi-asserted-by":"crossref","unstructured":"David L Raunig Lisa M McShane Gene Pennello Constantine Gatsonis Paul L Carson James T Voyvodic Richard L Wahl Brenda F Kurland Adam J Schwarz Mithat G\u00f6nen et al. 2015. Quantitative imaging biomarkers: a review of statistical methods for technical performance assessment. Statistical methods in medical research Vol. 24 1 (2015) 27--67.","DOI":"10.1177\/0962280214537344"},{"volume-title":"Modern statistical methods for HCI","author":"Robertson Judy","key":"e_1_2_2_48_1","unstructured":"Judy Robertson and Maurits Kaptein. 2016. Modern statistical methods for HCI. Vol. 6. Springer."},{"key":"e_1_2_2_49_1","doi-asserted-by":"publisher","DOI":"10.1049\/cit2.12081"},{"key":"e_1_2_2_50_1","volume-title":"Henriikka Vartiainen, Jarkko Suhonen, and Markku Tukiainen.","author":"Sanusi Ismaila Temitayo","year":"2022","unstructured":"Ismaila Temitayo Sanusi, Solomon Sunday Oyelere, Henriikka Vartiainen, Jarkko Suhonen, and Markku Tukiainen. 2022. A systematic review of teaching and learning machine learning in K-12 education. Education and Information Technologies (2022), 1--31."},{"key":"e_1_2_2_51_1","volume-title":"Saturation in qualitative research: exploring its conceptualization and operationalization. Quality & quantity","author":"Saunders Benjamin","year":"2018","unstructured":"Benjamin Saunders, Julius Sim, Tom Kingstone, Shula Baker, Jackie Waterfield, Bernadette Bartlam, Heather Burroughs, and Clare Jinks. 2018. Saturation in qualitative research: exploring its conceptualization and operationalization. Quality & quantity, Vol. 52 (2018), 1893--1907."},{"key":"e_1_2_2_52_1","unstructured":"Matthias Scheutz. [n. d.]. Ethical Aspects and Challenges for Interactive Task Learning. ( [n. d.])."},{"volume-title":"Guide to advanced empirical software engineering","author":"Shull Forrest","key":"e_1_2_2_53_1","unstructured":"Forrest Shull, Janice Singer, and Dag IK Sj\u00f8berg. 2007. Guide to advanced empirical software engineering. Springer."},{"key":"e_1_2_2_54_1","doi-asserted-by":"publisher","DOI":"10.1145\/3338283"},{"key":"e_1_2_2_55_1","doi-asserted-by":"publisher","DOI":"10.1145\/2531602.2531690"},{"key":"e_1_2_2_56_1","doi-asserted-by":"crossref","unstructured":"L\u00e1szl\u00f3 T\u00f3th Bal\u00e1zs Nagy D\u00e1vid Janth\u00f3 L\u00e1szl\u00f3 Vid\u00e1cs and Tibor Gyim\u00f3thy. 2019. Towards an accurate prediction of the question quality on stack overflow using a deep-learning-based nlp approach.. In ICSOFT. 631--639.","DOI":"10.5220\/0007971306310639"},{"key":"e_1_2_2_57_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2020.2968391"},{"key":"e_1_2_2_58_1","doi-asserted-by":"publisher","DOI":"10.1145\/3468264.3468558"},{"key":"e_1_2_2_59_1","doi-asserted-by":"publisher","DOI":"10.1145\/2876034.2876042"},{"key":"e_1_2_2_60_1","doi-asserted-by":"publisher","DOI":"10.1145\/3339474"},{"key":"e_1_2_2_61_1","doi-asserted-by":"publisher","DOI":"10.1109\/tse.2019.2906315"},{"key":"e_1_2_2_62_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.neucom.2017.01.026"},{"key":"e_1_2_2_63_1","first-page":"391","article-title":"Correcting Two-Sample ''z'' and ''t'' Tests for Correlation: An Alternative to One-Sample Tests on Difference Scores. Psicologica","volume":"33","author":"Zimmerman Donald W","year":"2012","unstructured":"Donald W Zimmerman. 2012. Correcting Two-Sample ''z'' and ''t'' Tests for Correlation: An Alternative to One-Sample Tests on Difference Scores. Psicologica: International Journal of Methodology and Experimental Psychology, Vol. 33, 2 (2012), 391--418.","journal-title":"International Journal of Methodology and Experimental Psychology"}],"container-title":["Proceedings of the ACM on Human-Computer Interaction"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3686990","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3686990","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,8,21]],"date-time":"2025-08-21T00:54:23Z","timestamp":1755737663000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3686990"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,11,7]]},"references-count":63,"journal-issue":{"issue":"CSCW2","published-print":{"date-parts":[[2024,11,7]]}},"alternative-id":["10.1145\/3686990"],"URL":"https:\/\/doi.org\/10.1145\/3686990","relation":{},"ISSN":["2573-0142"],"issn-type":[{"type":"electronic","value":"2573-0142"}],"subject":[],"published":{"date-parts":[[2024,11,7]]},"assertion":[{"value":"2024-11-08","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}