{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,21]],"date-time":"2026-07-21T14:44:38Z","timestamp":1784645078378,"version":"3.55.0"},"reference-count":68,"publisher":"Association for Computing Machinery (ACM)","issue":"1","license":[{"start":{"date-parts":[[2022,3,29]],"date-time":"2022-03-29T00:00:00Z","timestamp":1648512000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["Proc. ACM Interact. Mob. Wearable Ubiquitous Technol."],"published-print":{"date-parts":[[2022,3,29]]},"abstract":"<jats:p>A major bottleneck in training robust Human-Activity Recognition models (HAR) is the need for large-scale labeled sensor datasets. Because labeling large amounts of sensor data is an expensive task, unsupervised and semi-supervised learning techniques have emerged that can learn good features from the data without requiring any labels. In this paper, we extend this line of research and present a novel technique called Collaborative Self-Supervised Learning (ColloSSL) which leverages unlabeled data collected from multiple devices worn by a user to learn high-quality features of the data. A key insight that underpins the design of ColloSSL is that unlabeled sensor datasets simultaneously captured by multiple devices can be viewed as natural transformations of each other, and leveraged to generate a supervisory signal for representation learning. We present three technical innovations to extend conventional self-supervised learning algorithms to a multi-device setting: a Device Selection approach which selects positive and negative devices to enable contrastive learning, a Contrastive Sampling algorithm which samples positive and negative examples in a multi-device setting, and a loss function called Multi-view Contrastive Loss which extends standard contrastive loss to a multi-device setting. Our experimental results on three multi-device datasets show that ColloSSL outperforms both fully-supervised and semi-supervised learning techniques in majority of the experiment settings, resulting in an absolute increase of upto 7.9% in F1 score compared to the best performing baselines. We also show that ColloSSL outperforms the fully-supervised methods in a low-data regime, by just using one-tenth of the available labeled data in the best case.<\/jats:p>","DOI":"10.1145\/3517246","type":"journal-article","created":{"date-parts":[[2022,3,29]],"date-time":"2022-03-29T13:42:46Z","timestamp":1648561366000},"page":"1-28","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":109,"title":["ColloSSL"],"prefix":"10.1145","volume":"6","author":[{"given":"Yash","family":"Jain","sequence":"first","affiliation":[{"name":"Georgia Tech, United States of America and Nokia Bell Labs, United Kingdom"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Chi Ian","family":"Tang","sequence":"additional","affiliation":[{"name":"University of Cambridge, United Kingdom and Nokia Bell Labs, United Kingdom"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Chulhong","family":"Min","sequence":"additional","affiliation":[{"name":"Nokia Bell Labs, United Kingdom"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Fahim","family":"Kawsar","sequence":"additional","affiliation":[{"name":"Nokia Bell Labs, United Kingdom"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Akhil","family":"Mathur","sequence":"additional","affiliation":[{"name":"Nokia Bell Labs, United Kingdom"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2022,3,29]]},"reference":[{"key":"e_1_2_2_1_1","doi-asserted-by":"publisher","DOI":"10.1017\/S0373463307004560"},{"key":"e_1_2_2_2_1","first-page":"3","article-title":"A public domain dataset for human activity recognition using smartphones","volume":"3","author":"Anguita Davide","year":"2013","unstructured":"Davide Anguita, Alessandro Ghio, Luca Oneto, Xavier Parra, Jorge Luis Reyes-Ortiz, et al. 2013. A public domain dataset for human activity recognition using smartphones.. In Esann, Vol. 3. 3.","journal-title":"Esann"},{"key":"e_1_2_2_3_1","volume-title":"Learning representations by maximizing mutual information across views. arXiv preprint arXiv:1906.00910","author":"Bachman Philip","year":"2019","unstructured":"Philip Bachman, R Devon Hjelm, and William Buchwalter. 2019. Learning representations by maximizing mutual information across views. arXiv preprint arXiv:1906.00910 (2019)."},{"key":"e_1_2_2_4_1","volume-title":"Proceedings of ICML Workshop on Unsupervised and Transfer Learning (Proceedings of Machine Learning Research","volume":"49","author":"Baldi Pierre","year":"2012","unstructured":"Pierre Baldi. 2012. Autoencoders, Unsupervised Learning, and Deep Architectures. In Proceedings of ICML Workshop on Unsupervised and Transfer Learning (Proceedings of Machine Learning Research, Vol. 27), Isabelle Guyon, Gideon Dror, Vincent Lemaire, Graham Taylor, and Daniel Silver (Eds.). PMLR, Bellevue, Washington, USA, 37--49. http:\/\/proceedings.mlr.press\/v27\/baldi12a.html"},{"key":"e_1_2_2_5_1","doi-asserted-by":"publisher","DOI":"10.1145\/2499621"},{"key":"e_1_2_2_6_1","doi-asserted-by":"crossref","unstructured":"Mathilde Caron Hugo Touvron Ishan Misra Herv\u00e9 J\u00e9gou Julien Mairal Piotr Bojanowski and Armand Joulin. 2021. Emerging Properties in Self-Supervised Vision Transformers. arXiv:2104.14294 [cs.CV]","DOI":"10.1109\/ICCV48922.2021.00951"},{"key":"e_1_2_2_7_1","doi-asserted-by":"publisher","DOI":"10.1145\/3380985"},{"key":"e_1_2_2_8_1","volume-title":"International conference on machine learning. PMLR, 1597--1607","author":"Chen Ting","year":"2020","unstructured":"Ting Chen, Simon Kornblith, Mohammad Norouzi, and Geoffrey Hinton. 2020. A simple framework for contrastive learning of visual representations. In International conference on machine learning. PMLR, 1597--1607."},{"key":"e_1_2_2_9_1","doi-asserted-by":"publisher","DOI":"10.1145\/3478119"},{"key":"e_1_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.01549"},{"key":"e_1_2_2_11_1","doi-asserted-by":"publisher","DOI":"10.1145\/3442381.3449903"},{"key":"e_1_2_2_12_1","volume-title":"Romit Roy Choudhury, and Srihari Nelakuditi","author":"Dey Sanorita","year":"2014","unstructured":"Sanorita Dey, Nirupam Roy, Wenyuan Xu, Romit Roy Choudhury, and Srihari Nelakuditi. 2014. AccelPrint: Imperfections of Accelerometers Make Smartphones Trackable.. In NDSS. Citeseer."},{"key":"e_1_2_2_13_1","doi-asserted-by":"publisher","DOI":"10.1007\/s00779-010-0293-9"},{"key":"e_1_2_2_14_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIM.2008.2006137"},{"key":"e_1_2_2_15_1","doi-asserted-by":"publisher","DOI":"10.1145\/3191743"},{"key":"e_1_2_2_16_1","volume-title":"A kernel method for the two-sample problem. arXiv preprint arXiv:0805.2368","author":"Gretton Arthur","year":"2008","unstructured":"Arthur Gretton, Karsten Borgwardt, Malte J Rasch, Bernhard Scholkopf, and Alexander J Smola. 2008. A kernel method for the two-sample problem. arXiv preprint arXiv:0805.2368 (2008)."},{"key":"e_1_2_2_17_1","volume-title":"Zhaohan Daniel Guo, Mohammad Gheshlaghi Azar, et al.","author":"Grill Jean-Bastien","year":"2020","unstructured":"Jean-Bastien Grill, Florian Strub, Florent Altch\u00e9, Corentin Tallec, Pierre H Richemond, Elena Buchatskaya, Carl Doersch, Bernardo Avila Pires, Zhaohan Daniel Guo, Mohammad Gheshlaghi Azar, et al. 2020. Bootstrap your own latent: A new approach to self-supervised learning. arXiv preprint arXiv:2006.07733 (2020)."},{"key":"e_1_2_2_18_1","doi-asserted-by":"publisher","DOI":"10.1145\/3090076"},{"key":"e_1_2_2_19_1","volume-title":"convolutional, and recurrent models for human activity recognition using wearables. arXiv preprint arXiv:1604.08880","author":"Hammerla Nils Y","year":"2016","unstructured":"Nils Y Hammerla, Shane Halloran, and Thomas Pl\u00f6tz. 2016. Deep, convolutional, and recurrent models for human activity recognition using wearables. arXiv preprint arXiv:1604.08880 (2016)."},{"key":"e_1_2_2_20_1","doi-asserted-by":"publisher","DOI":"10.1145\/2493988.2494353"},{"key":"e_1_2_2_21_1","doi-asserted-by":"publisher","DOI":"10.1145\/3410531.3414306"},{"key":"e_1_2_2_22_1","doi-asserted-by":"publisher","DOI":"10.1145\/3463506"},{"key":"e_1_2_2_23_1","volume-title":"Hsiao-Yu Fish Tung, and Katerina Fragkiadaki","author":"Harley Adam W.","year":"2020","unstructured":"Adam W. Harley, Shrinidhi K. Lakshmikanth, Fangyu Li, Xian Zhou, Hsiao-Yu Fish Tung, and Katerina Fragkiadaki. 2020. Learning from Unlabelled Videos Using Contrastive Predictive Neural 3D Mapping. arXiv:1906.03764 [cs.CV]"},{"key":"e_1_2_2_24_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.jnca.2019.06.016"},{"key":"e_1_2_2_25_1","unstructured":"Kaiming He Haoqi Fan Yuxin Wu Saining Xie and Ross Girshick. 2020. Momentum Contrast for Unsupervised Visual Representation Learning. arXiv:1911.05722 [cs.CV]"},{"key":"e_1_2_2_26_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2020.2992393"},{"key":"e_1_2_2_27_1","doi-asserted-by":"crossref","unstructured":"Seungwoo Kang Youngki Lee Chulhong Min Younghyun Ju Taiwoo Park Jinwon Lee Yunseok Rhee and Junehwa Song. 2010. Orchestrator: An active resource orchestration framework for mobile context monitoring in sensor-rich mobile environments. In 2010 ieee international conference on pervasive computing and communications (percom). IEEE 135--144.","DOI":"10.1109\/PERCOM.2010.5466982"},{"key":"e_1_2_2_28_1","doi-asserted-by":"publisher","DOI":"10.1145\/2070942.2070968"},{"key":"e_1_2_2_29_1","unstructured":"Prannay Khosla Piotr Teterwak Chen Wang Aaron Sarna Yonglong Tian Phillip Isola Aaron Maschinot Ce Liu and Dilip Krishnan. 2021. Supervised Contrastive Learning. arXiv:2004.11362 [cs.LG]"},{"key":"e_1_2_2_30_1","doi-asserted-by":"publisher","DOI":"10.1145\/1409635.1409639"},{"key":"e_1_2_2_31_1","volume-title":"Sang Min Yoon, and Heeryon Cho","author":"Lee Song-Mi","year":"2017","unstructured":"Song-Mi Lee, Sang Min Yoon, and Heeryon Cho. 2017. Human activity recognition from accelerometer data using Convolutional Neural Network. In 2017 ieee international conference on big data and smart computing (bigcomp). IEEE, 131--134."},{"key":"e_1_2_2_32_1","doi-asserted-by":"publisher","DOI":"10.1145\/2994374.2994388"},{"key":"e_1_2_2_33_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPSN.2018.00048"},{"key":"e_1_2_2_34_1","volume-title":"A survey on bias and fairness in machine learning. arXiv preprint arXiv:1908.09635","author":"Mehrabi Ninareh","year":"2019","unstructured":"Ninareh Mehrabi, Fred Morstatter, Nripsuta Saxena, Kristina Lerman, and Aram Galstyan. 2019. A survey on bias and fairness in machine learning. arXiv preprint arXiv:1908.09635 (2019)."},{"key":"e_1_2_2_35_1","doi-asserted-by":"publisher","DOI":"10.3390\/app7101101"},{"key":"e_1_2_2_36_1","doi-asserted-by":"publisher","DOI":"10.1145\/3356250.3360043"},{"key":"e_1_2_2_37_1","doi-asserted-by":"publisher","DOI":"10.3390\/s16010115"},{"key":"e_1_2_2_38_1","volume-title":"Specaugment: A simple data augmentation method for automatic speech recognition. arXiv preprint arXiv:1904.08779","author":"Park Daniel S","year":"2019","unstructured":"Daniel S Park, William Chan, Yu Zhang, Chung-Cheng Chiu, Barret Zoph, Ekin D Cubuk, and Quoc V Le. 2019. Specaugment: A simple data augmentation method for automatic speech recognition. arXiv preprint arXiv:1904.08779 (2019)."},{"key":"e_1_2_2_39_1","doi-asserted-by":"publisher","DOI":"10.1145\/3214277"},{"key":"e_1_2_2_40_1","doi-asserted-by":"publisher","DOI":"10.1145\/3459666"},{"key":"e_1_2_2_41_1","unstructured":"Thomas Pl\u00f6tz Nils Y Hammerla and Patrick L Olivier. 2011. Feature learning for activity recognition in ubiquitous computing. In Twenty-second international joint conference on artificial intelligence."},{"key":"e_1_2_2_42_1","doi-asserted-by":"publisher","DOI":"10.1115\/1.4034419"},{"key":"e_1_2_2_43_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISWC.2012.13"},{"key":"e_1_2_2_44_1","doi-asserted-by":"publisher","DOI":"10.1109\/INSS.2010.5573462"},{"key":"e_1_2_2_45_1","volume-title":"Human activity recognition with smartphone sensors using deep learning neural networks. Expert systems with applications 59","author":"Ronao Charissa Ann","year":"2016","unstructured":"Charissa Ann Ronao and Sung-Bae Cho. 2016. Human activity recognition with smartphone sensors using deep learning neural networks. Expert systems with applications 59 (2016), 235--244."},{"key":"e_1_2_2_46_1","doi-asserted-by":"publisher","DOI":"10.1145\/3328932"},{"key":"e_1_2_2_47_1","doi-asserted-by":"publisher","DOI":"10.1109\/JIOT.2020.3009358"},{"key":"e_1_2_2_48_1","volume-title":"Sense and Learn: Self-Supervision for Omnipresent Sensors. arXiv preprint arXiv:2009.13233","author":"Saeed Aaqib","year":"2020","unstructured":"Aaqib Saeed, Victor Ungureanu, and Beat Gfeller. 2020. Sense and Learn: Self-Supervision for Omnipresent Sensors. arXiv preprint arXiv:2009.13233 (2020)."},{"key":"e_1_2_2_49_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICSRS.2017.8272822"},{"key":"e_1_2_2_50_1","volume-title":"Self-supervised ecg representation learning for emotion recognition. arXiv preprint arXiv:2002.03898","author":"Sarkar Pritam","year":"2020","unstructured":"Pritam Sarkar and Ali Etemad. 2020. Self-supervised ecg representation learning for emotion recognition. arXiv preprint arXiv:2002.03898 (2020)."},{"key":"e_1_2_2_51_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2018.8462891"},{"key":"e_1_2_2_52_1","volume-title":"Federated Self-Supervised Contrastive Learning via Ensemble Similarity Distillation. arXiv preprint arXiv:2109.14611","author":"Shi Haizhou","year":"2021","unstructured":"Haizhou Shi, Youcai Zhang, Zijin Shen, Siliang Tang, Yaqian Li, Yandong Guo, and Yueting Zhuang. 2021. Federated Self-Supervised Contrastive Learning via Ensemble Similarity Distillation. arXiv preprint arXiv:2109.14611 (2021)."},{"key":"e_1_2_2_53_1","volume-title":"Deep inside convolutional networks: Visualising image classification models and saliency maps. arXiv preprint arXiv:1312.6034","author":"Simonyan Karen","year":"2013","unstructured":"Karen Simonyan, Andrea Vedaldi, and Andrew Zisserman. 2013. Deep inside convolutional networks: Visualising image classification models and saliency maps. arXiv preprint arXiv:1312.6034 (2013)."},{"key":"e_1_2_2_54_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISWC.2008.4911590"},{"key":"e_1_2_2_55_1","doi-asserted-by":"publisher","DOI":"10.1145\/2809695.2809718"},{"key":"e_1_2_2_56_1","doi-asserted-by":"publisher","DOI":"10.1109\/PERCOM.2016.7456521"},{"key":"e_1_2_2_57_1","volume-title":"SelfHAR: Improving Human Activity Recognition through Self-training with Unlabeled Data. arXiv preprint arXiv: 2102.06073","author":"Tang Chi Ian","year":"2021","unstructured":"Chi Ian Tang, Ignacio Perez-Pozuelo, Dimitris Spathis, Soren Brage, Nick Wareham, and Cecilia Mascolo. 2021. SelfHAR: Improving Human Activity Recognition through Self-training with Unlabeled Data. arXiv preprint arXiv: 2102.06073 (2021)."},{"key":"e_1_2_2_58_1","volume-title":"Exploring Contrastive Learning in Human Activity Recognition for Healthcare. arXiv preprint arXiv:2011.11542","author":"Tang Chi Ian","year":"2020","unstructured":"Chi Ian Tang, Ignacio Perez-Pozuelo, Dimitris Spathis, and Cecilia Mascolo. 2020. Exploring Contrastive Learning in Human Activity Recognition for Healthcare. arXiv preprint arXiv:2011.11542 (2020)."},{"key":"e_1_2_2_59_1","volume-title":"Domain adaptation in computer vision applications","author":"Tommasi Tatiana","unstructured":"Tatiana Tommasi, Novi Patricia, Barbara Caputo, and Tinne Tuytelaars. 2017. A deeper look at dataset bias. In Domain adaptation in computer vision applications. Springer, 37--55."},{"key":"e_1_2_2_60_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2011.5995347"},{"key":"e_1_2_2_61_1","doi-asserted-by":"publisher","DOI":"10.1145\/3161192"},{"key":"e_1_2_2_62_1","article-title":"Visualizing data using t-SNE","volume":"9","author":"der Maaten Laurens Van","year":"2008","unstructured":"Laurens Van der Maaten and Geoffrey Hinton. 2008. Visualizing data using t-SNE. Journal of machine learning research 9, 11 (2008).","journal-title":"Journal of machine learning research"},{"key":"e_1_2_2_63_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10994-019-05855-6"},{"key":"e_1_2_2_64_1","doi-asserted-by":"publisher","DOI":"10.5555\/2832747.2832806"},{"key":"e_1_2_2_65_1","doi-asserted-by":"publisher","DOI":"10.1145\/3038912.3052577"},{"key":"e_1_2_2_66_1","doi-asserted-by":"publisher","DOI":"10.1145\/3212725.3212729"},{"key":"e_1_2_2_67_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-540-77690-1_2"},{"key":"e_1_2_2_68_1","unstructured":"Xiaojin Jerry Zhu. 2005. Semi-supervised learning literature survey. Technical Report. University of Wisconsin-Madison Department of Computer Sciences."}],"container-title":["Proceedings of the ACM on Interactive, Mobile, Wearable and Ubiquitous Technologies"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3517246","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3517246","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,7,14]],"date-time":"2025-07-14T04:24:50Z","timestamp":1752467090000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3517246"}},"subtitle":["Collaborative Self-Supervised Learning for Human Activity Recognition"],"short-title":[],"issued":{"date-parts":[[2022,3,29]]},"references-count":68,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2022,3,29]]}},"alternative-id":["10.1145\/3517246"],"URL":"https:\/\/doi.org\/10.1145\/3517246","relation":{},"ISSN":["2474-9567"],"issn-type":[{"value":"2474-9567","type":"electronic"}],"subject":[],"published":{"date-parts":[[2022,3,29]]},"assertion":[{"value":"2022-03-29","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}