{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,12,29]],"date-time":"2025-12-29T22:18:23Z","timestamp":1767046703478,"version":"3.41.0"},"reference-count":52,"publisher":"Association for Computing Machinery (ACM)","issue":"1","license":[{"start":{"date-parts":[[2011,1,1]],"date-time":"2011-01-01T00:00:00Z","timestamp":1293840000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/100000145","name":"Division of Information and Intelligent Systems","doi-asserted-by":"publisher","award":["IIS-0433637IIS-0845683"],"award-info":[{"award-number":["IIS-0433637IIS-0845683"]}],"id":[{"id":"10.13039\/100000145","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Intell. Syst. Technol."],"published-print":{"date-parts":[[2011,1]]},"abstract":"<jats:p>\n            New technologies have made it possible to collect information about social networks as they are acted and observed\n            <jats:italic>in the wild<\/jats:italic>\n            , instead of as they are reported in retrospective surveys. These technologies offer opportunities to address many new research questions: How can meaningful information about social interaction be extracted from automatically recorded raw data on human behavior? What can we learn about social networks from such fine-grained behavioral data? And how can all of this be done while protecting privacy? With the goal of addressing these questions, this article presents new methods for inferring colocation and conversation networks from privacy-sensitive audio. These methods are applied in a study of face-to-face interactions among 24 students in a graduate school cohort during an academic year. The resulting analysis shows that networks derived from colocation and conversation inferences are quite different. This distinction can inform future research in computational social science, especially work that only measures colocation or employs colocation data as a proxy for conversation networks.\n          <\/jats:p>","DOI":"10.1145\/1889681.1889688","type":"journal-article","created":{"date-parts":[[2012,10,12]],"date-time":"2012-10-12T20:56:02Z","timestamp":1350075362000},"page":"1-41","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":48,"title":["Inferring colocation and conversation networks from privacy-sensitive audio with implications for computational social science"],"prefix":"10.1145","volume":"2","author":[{"given":"Danny","family":"Wyatt","sequence":"first","affiliation":[{"name":"University of Washington, Seattle, WA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Tanzeem","family":"Choudhury","sequence":"additional","affiliation":[{"name":"Dartmouth College, Hanover, NH"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jeff","family":"Bilmes","sequence":"additional","affiliation":[{"name":"University of Washington, Seattle, WA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"James A.","family":"Kitts","sequence":"additional","affiliation":[{"name":"Columbia University, New York, NY"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2011,1,24]]},"reference":[{"key":"e_1_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2004.1326058"},{"key":"e_1_2_1_2_1","volume-title":"Proceedings of the International Conference on Spoken Language Processing (ICSLP).","author":"Ang J.","year":"2002","unstructured":"Ang , J. 2002 . Prosody-based automatic detection of annoyance and frustration in human-computer dialog . In Proceedings of the International Conference on Spoken Language Processing (ICSLP). Ang, J. 2002. Prosody-based automatic detection of annoyance and frustration in human-computer dialog. In Proceedings of the International Conference on Spoken Language Processing (ICSLP)."},{"key":"e_1_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2003.1198906"},{"volume-title":"Proceeding of the ISCA Tutorial and Research Workshop on Speech and Emotion.","author":"Batliner A.","key":"e_1_2_1_6_1","unstructured":"Batliner , A. , Fisher , K. , Huber , R. , Spilker , J. , and N\u00f6th , E . 2000. Desperately seeking emotions or: actors, wizards and human beings . In Proceeding of the ISCA Tutorial and Research Workshop on Speech and Emotion. Batliner, A., Fisher, K., Huber, R., Spilker, J., and N\u00f6th, E. 2000. Desperately seeking emotions or: actors, wizards and human beings. In Proceeding of the ISCA Tutorial and Research Workshop on Speech and Emotion."},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1177\/1461444804041438"},{"key":"e_1_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1111\/j.1468-2958.1977.tb00591.x"},{"key":"e_1_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1016\/0378-8733(79)90014-5"},{"key":"e_1_2_1_10_1","first-page":"30","article-title":"Informant accuracy in social network data V: An experimental attempt to predict actual communication from recall data. Social Sci","volume":"11","author":"Bernard H. R.","year":"1982","unstructured":"Bernard , H. R. , Killworth , P. D. , and Sailer , L. 1982 . Informant accuracy in social network data V: An experimental attempt to predict actual communication from recall data. Social Sci . Resear. 11 , 30 -- 66 . Bernard, H. R., Killworth, P. D., and Sailer, L. 1982. Informant accuracy in social network data V: An experimental attempt to predict actual communication from recall data. Social Sci. Resear. 11, 30--66.","journal-title":"Resear."},{"volume-title":"Department of Electrical Engineering","author":"Bilmes J.","key":"e_1_2_1_11_1","unstructured":"Bilmes , J. 2004. On soft evidence in bayesian networks. Tech. rep. 16 , Department of Electrical Engineering , University of Washingon. Bilmes, J. 2004. On soft evidence in bayesian networks. Tech. rep. 16, Department of Electrical Engineering, University of Washingon."},{"key":"e_1_2_1_13_1","volume-title":"Proceedings of the Annual Conference on Language Resources and Evaluation (LREC).","author":"Campbell N.","year":"2002","unstructured":"Campbell , N. 2002 . The recording of emotional speech: JST\/CREST database research . In Proceedings of the Annual Conference on Language Resources and Evaluation (LREC). Campbell, N. 2002. The recording of emotional speech: JST\/CREST database research. In Proceedings of the Annual Conference on Language Resources and Evaluation (LREC)."},{"volume-title":"Proceedings of the International Conference on Wearable Computing.","author":"Choudhury T.","key":"e_1_2_1_15_1","unstructured":"Choudhury , T. and Pentland , A. S . 2003. Sensing and modeling human networks using the sociometer . In Proceedings of the International Conference on Wearable Computing. Choudhury, T. and Pentland, A. S. 2003. Sensing and modeling human networks using the sociometer. In Proceedings of the International Conference on Wearable Computing."},{"key":"e_1_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/WACV.2008.4544042"},{"key":"e_1_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1016\/0378-8733(94)90003-5"},{"key":"e_1_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.2307\/2093295"},{"volume-title":"Proceedings of the International Conference on Spoken Language Processing (ICSLP).","author":"Dellaert F.","key":"e_1_2_1_19_1","unstructured":"Dellaert , F. , Polzin , T. , and Waibel , A . 1996. Recognizing emotion in speech . In Proceedings of the International Conference on Spoken Language Processing (ICSLP). Dellaert, F., Polzin, T., and Waibel, A. 1996. Recognizing emotion in speech. In Proceedings of the International Conference on Spoken Language Processing (ICSLP)."},{"volume-title":"Proceedings of the IEEE Workshop on Multimedia Signal Processing.","author":"Dielmann A.","key":"e_1_2_1_20_1","unstructured":"Dielmann , A. and Renals , S . 2004. Multi-stream segmentation of meetings . In Proceedings of the IEEE Workshop on Multimedia Signal Processing. Dielmann, A. and Renals, S. 2004. Multi-stream segmentation of meetings. In Proceedings of the IEEE Workshop on Multimedia Signal Processing."},{"key":"e_1_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1016\/S0167-6393(02)00070-5"},{"volume-title":"Proceedings of the ISCA Tutorial and Research Workshop on Speech and Emotion.","author":"Douglas-Cowie E.","key":"e_1_2_1_23_1","unstructured":"Douglas-Cowie , E. , Cowie , R. , and Schroeder , M . 2000. A new emotion database: considerations, sources and scope . In Proceedings of the ISCA Tutorial and Research Workshop on Speech and Emotion. Douglas-Cowie, E., Cowie, R., and Schroeder, M. 2000. A new emotion database: considerations, sources and scope. In Proceedings of the ISCA Tutorial and Research Workshop on Speech and Emotion."},{"key":"e_1_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1007\/s00779-005-0046-3"},{"volume-title":"Proceedings of Robotics: Science and Systems.","author":"Ferris B.","key":"e_1_2_1_25_1","unstructured":"Ferris , B. , Haehnel , D. , and Fox , D . 2006. Gaussian processes for signal strength-based location estimation . In Proceedings of Robotics: Science and Systems. Ferris, B., Haehnel, D., and Fox, D. 2006. Gaussian processes for signal strength-based location estimation. In Proceedings of Robotics: Science and Systems."},{"key":"e_1_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.2307\/2786941"},{"key":"e_1_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1525\/aa.1987.89.2.02a00020"},{"volume-title":"Proceedings of the International Conference on Acoustics, Speech, and Signal Processing (ICASSP).","author":"Gatica-Perez D.","key":"e_1_2_1_28_1","unstructured":"Gatica-Perez , D. , McCowan , I. , Zhang , D. , and Bengio , S . 2005. Detecting group interest-level in meetings . In Proceedings of the International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Gatica-Perez, D., McCowan, I., Zhang, D., and Bengio, S. 2005. Detecting group interest-level in meetings. In Proceedings of the International Conference on Acoustics, Speech, and Signal Processing (ICASSP)."},{"key":"e_1_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1353\/dem.0.0045"},{"key":"e_1_2_1_30_1","doi-asserted-by":"crossref","unstructured":"Gray R. M. and Davisson L. D. 2004. An Introduction to Statistical Signal Processing. Cambridge University Press.  Gray R. M. and Davisson L. D. 2004. An Introduction to Statistical Signal Processing. Cambridge University Press.","DOI":"10.1017\/CBO9780511801372"},{"volume-title":"Proceedings of the International Congress of Phonetic Sciences.","author":"Greasley P.","key":"e_1_2_1_31_1","unstructured":"Greasley , P. , Setter , J. , Waterman , M. , Sherrard , C. , Roach , P. , Arnfield , S. , and Horton , D . 1995. Representation of prosodic and emotional features in a spoken language database . In Proceedings of the International Congress of Phonetic Sciences. Greasley, P., Setter, J., Waterman, M., Sherrard, C., Roach, P., Arnfield, S., and Horton, D. 1995. Representation of prosodic and emotional features in a spoken language database. In Proceedings of the International Congress of Phonetic Sciences."},{"key":"e_1_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1177\/0261927X91103003"},{"key":"e_1_2_1_33_1","doi-asserted-by":"crossref","unstructured":"Holland P. W. and Leinhardt S. 1975. The statistical analysis of local structure in social networks. In Sociological Methodology Jossey-Bass 1--45.  Holland P. W. and Leinhardt S. 1975. The statistical analysis of local structure in social networks. In Sociological Methodology Jossey-Bass 1--45.","DOI":"10.2307\/270703"},{"key":"e_1_2_1_34_1","doi-asserted-by":"publisher","DOI":"10.1023\/A:1013849922756"},{"key":"e_1_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.2189\/asqu.52.4.558"},{"key":"e_1_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.18637\/jss.v028.c01"},{"key":"e_1_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.17730\/humo.35.3.10215j2m359266n2"},{"key":"e_1_2_1_38_1","doi-asserted-by":"publisher","DOI":"10.1016\/0378-8733(79)90009-1"},{"key":"e_1_2_1_39_1","doi-asserted-by":"publisher","DOI":"10.1126\/science.1116869"},{"key":"e_1_2_1_40_1","doi-asserted-by":"publisher","DOI":"10.1016\/S0378-8733(97)00006-3"},{"key":"e_1_2_1_41_1","doi-asserted-by":"publisher","DOI":"10.1145\/1367497.1367591"},{"volume-title":"Proceedings of the International Joint Conference on Artificial Intelligence (IJCAI).","author":"Lester J.","key":"e_1_2_1_42_1","unstructured":"Lester , J. , Choudhury , T. , Kern , N. , Borriello , G. , and Hannaford , B . 2005. A hybrid discriminative-generative approach for modeling human activities . In Proceedings of the International Joint Conference on Artificial Intelligence (IJCAI). Lester, J., Choudhury, T., Kern, N., Borriello, G., and Hannaford, B. 2005. A hybrid discriminative-generative approach for modeling human activities. In Proceedings of the International Joint Conference on Artificial Intelligence (IJCAI)."},{"volume-title":"Proceedings of the International Joint Conference on Artificial Intelligence (IJCAI).","author":"Lian C.","key":"e_1_2_1_43_1","unstructured":"Lian , C. and Hsu , J . 2009. Probabilistic models for concurrent chatting activity recognition . In Proceedings of the International Joint Conference on Artificial Intelligence (IJCAI). Lian, C. and Hsu, J. 2009. Probabilistic models for concurrent chatting activity recognition. In Proceedings of the International Joint Conference on Artificial Intelligence (IJCAI)."},{"volume-title":"Proceedings of the International Conference on Acoustics, Speech, and Signal Processing (ICASSP).","author":"McCowan I.","key":"e_1_2_1_44_1","unstructured":"McCowan , I. , Bengio , S. , Gatica-Perez , D. , Lathoud , G. , Monay , F. , Moore , D. , Wellner , P. , and Bourlard , H . 2003. Modeling human interaction in meetings . In Proceedings of the International Conference on Acoustics, Speech, and Signal Processing (ICASSP). McCowan, I., Bengio, S., Gatica-Perez, D., Lathoud, G., Monay, F., Moore, D., Wellner, P., and Bourlard, H. 2003. Modeling human interaction in meetings. In Proceedings of the International Conference on Acoustics, Speech, and Signal Processing (ICASSP)."},{"key":"e_1_2_1_45_1","unstructured":"NIST. 2009. NIST rich transcription evaluations. http:\/\/www.itl.nist.gov\/iad\/mig\/tests\/rt\/2009\/index.html.  NIST. 2009. NIST rich transcription evaluations. http:\/\/www.itl.nist.gov\/iad\/mig\/tests\/rt\/2009\/index.html."},{"key":"e_1_2_1_46_1","doi-asserted-by":"publisher","DOI":"10.1088\/1367-2630\/9\/6\/179"},{"key":"e_1_2_1_47_1","doi-asserted-by":"publisher","DOI":"10.1038\/nature05670"},{"volume-title":"Discrete-Time Speech Signal Processing: Principles and Practice","author":"Quatieri T.","key":"e_1_2_1_48_1","unstructured":"Quatieri , T. 2001. Discrete-Time Speech Signal Processing: Principles and Practice . Prentice Hall . Quatieri, T. 2001. Discrete-Time Speech Signal Processing: Principles and Practice. Prentice Hall."},{"key":"e_1_2_1_49_1","doi-asserted-by":"publisher","DOI":"10.1109\/TASSP.1977.1162905"},{"key":"e_1_2_1_50_1","doi-asserted-by":"publisher","DOI":"10.1109\/5.18626"},{"volume-title":"Proceedings of the International Conference on Acoustics, Speech, and Signal Processing (ICASSP).","author":"Reynolds D. A.","key":"e_1_2_1_51_1","unstructured":"Reynolds , D. A. and Torres-Carrasquillo , P . 2005. Approaches and applications of audio diarization . In Proceedings of the International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Reynolds, D. A. and Torres-Carrasquillo, P. 2005. Approaches and applications of audio diarization. In Proceedings of the International Conference on Acoustics, Speech, and Signal Processing (ICASSP)."},{"key":"e_1_2_1_52_1","doi-asserted-by":"publisher","DOI":"10.1103\/PhysRevE.75.027105"},{"volume-title":"Proceedings of the International Conference on Acoustics, Speech, and Signal Processing (ICASSP).","author":"Schuller B.","key":"e_1_2_1_53_1","unstructured":"Schuller , B. , Rigoll , G. , and Lang , M . 2004. Speech emotion recognition combining acoustic features and linguistic information in a hybrid support vector machine-belief network architecture . In Proceedings of the International Conference on Acoustics, Speech, and Signal Processing (ICASSP). Schuller, B., Rigoll, G., and Lang, M. 2004. Speech emotion recognition combining acoustic features and linguistic information in a hybrid support vector machine-belief network architecture. In Proceedings of the International Conference on Acoustics, Speech, and Signal Processing (ICASSP)."},{"key":"e_1_2_1_54_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2009.4960543"},{"key":"e_1_2_1_55_1","doi-asserted-by":"publisher","DOI":"10.1038\/30918"},{"key":"e_1_2_1_56_1","volume-title":"Tech. Rep. 2007-069, MERL.","author":"Wren C. R.","year":"2007","unstructured":"Wren , C. R. , Ivanov , Y. A. , Leigh , D. , and Westhues , J . 2007 . The MERL motion detector dataset. Tech. Rep. 2007-069, MERL. Wren, C. R., Ivanov, Y. A., Leigh, D., and Westhues, J. 2007. The MERL motion detector dataset. Tech. Rep. 2007-069, MERL."},{"volume-title":"Proceedings of Interspeech.","author":"Wyatt D.","key":"e_1_2_1_57_1","unstructured":"Wyatt , D. , Choudhury , T. , and Bilmes , J . 2007. Conversation detection and speaker segmentation in privacy-sensitive situated speech data . In Proceedings of Interspeech. Wyatt, D., Choudhury, T., and Bilmes, J. 2007. Conversation detection and speaker segmentation in privacy-sensitive situated speech data. In Proceedings of Interspeech."}],"container-title":["ACM Transactions on Intelligent Systems and Technology"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/1889681.1889688","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/1889681.1889688","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T10:52:17Z","timestamp":1750243937000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/1889681.1889688"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2011,1]]},"references-count":52,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2011,1]]}},"alternative-id":["10.1145\/1889681.1889688"],"URL":"https:\/\/doi.org\/10.1145\/1889681.1889688","relation":{},"ISSN":["2157-6904","2157-6912"],"issn-type":[{"type":"print","value":"2157-6904"},{"type":"electronic","value":"2157-6912"}],"subject":[],"published":{"date-parts":[[2011,1]]},"assertion":[{"value":"2010-08-01","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2010-10-01","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2011-01-24","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}