{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,10,17]],"date-time":"2025-10-17T14:22:46Z","timestamp":1760710966755,"version":"3.41.0"},"reference-count":54,"publisher":"Association for Computing Machinery (ACM)","issue":"2s","license":[{"start":{"date-parts":[[2022,6,30]],"date-time":"2022-06-30T00:00:00Z","timestamp":1656547200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"Institute of Information & Communications Technology Planning & Evaluation"},{"name":"Korea government","award":["2020-0-00062"],"award-info":[{"award-number":["2020-0-00062"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Multimedia Comput. Commun. Appl."],"published-print":{"date-parts":[[2022,6,30]]},"abstract":"<jats:p>\n            Deep Learning models\u2019 performance strongly correlate with availability of annotated data; however, massive data labeling is laborious, expensive, and error-prone when performed by human experts.\n            <jats:bold>Active Learning (AL)<\/jats:bold>\n            effectively handles this challenge by selecting the uncertain samples from unlabeled data collection, but the existing AL approaches involve repetitive human feedback for labeling uncertain samples, thus rendering these techniques infeasible to be deployed in industry related real-world applications. In the proposed\n            <jats:bold>Proxy Model based Active Learning technique (PMAL)<\/jats:bold>\n            , this issue is addressed by replacing human oracle with a deep learning model, where human expertise is reduced to label only two small subsets of data for training proxy model and initializing the AL loop. In the PMAL technique, firstly, proxy model is trained with a small subset of labeled data, which subsequently acts as an oracle for annotating uncertain samples. Secondly, active model's training, uncertain samples extraction via uncertainty sampling, and annotation through proxy model is carried out until predefined iterations to achieve higher accuracy and labeled data. Finally, the active model is evaluated using testing data to verify the effectiveness of our technique for practical applications. The correct annotations by the proxy model are ensured by employing the potentials of explainable artificial intelligence. Similarly, emerging vision transformer is used as an active model to achieve maximum accuracy. Experimental results reveal that the proposed method outperforms the state-of-the-art in terms of minimum labeled data usage and improves the accuracy with 2.2%, 2.6%, and 1.35% on Caltech-101, Caltech-256, and CIFAR-10 datasets, respectively. Since the proposed technique offers a highly reasonable solution to exploit huge multimedia data, it can be widely used in different evolutionary industrial domains.\n          <\/jats:p>","DOI":"10.1145\/3534932","type":"journal-article","created":{"date-parts":[[2022,6,21]],"date-time":"2022-06-21T11:03:15Z","timestamp":1655809395000},"page":"1-18","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":5,"title":["PMAL: A Proxy Model Active Learning Approach for Vision Based Industrial Applications"],"prefix":"10.1145","volume":"18","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-7551-6082","authenticated-orcid":false,"given":"Abbas","family":"Khan","sequence":"first","affiliation":[{"name":"Sejong University, Seoul, Republic of Korea"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-8201-7372","authenticated-orcid":false,"given":"Ijaz Ul","family":"Haq","sequence":"additional","affiliation":[{"name":"Sejong University, Seoul, Republic of Korea"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-4861-8347","authenticated-orcid":false,"given":"Tanveer","family":"Hussain","sequence":"additional","affiliation":[{"name":"Sejong University, Seoul, Republic of Korea"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-5302-1150","authenticated-orcid":false,"given":"Khan","family":"Muhammad","sequence":"additional","affiliation":[{"name":"Visual Analytics for Knowledge Laboratory (VIS2KNOW Lab), Department of Applied Artificial Intelligence, School of Convergence, College of Computing and Informatics, Sungkyunkwan University, Seoul, Republic of Korea"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-9279-401X","authenticated-orcid":false,"given":"Mohammad","family":"Hijji","sequence":"additional","affiliation":[{"name":"Faculty of Computers &amp; Information Technology, Computer Science Department, University of Tabuk, Tabuk, Saudi Arabia"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-5646-0338","authenticated-orcid":false,"given":"Muhammad","family":"Sajjad","sequence":"additional","affiliation":[{"name":"Digital Image Processing Laboratory, Department of Computer Science, Islamia College Peshawar, Peshawar, Pakistan"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-3886-4309","authenticated-orcid":false,"given":"Victor Hugo C.","family":"De Albuquerque","sequence":"additional","affiliation":[{"name":"Department of Teleinformatics Engineering, Federal University of Cear\u00e1, Fortaleza, Fortaleza\/CE, Brazil"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-6678-7788","authenticated-orcid":false,"given":"Sung Wook","family":"Baik","sequence":"additional","affiliation":[{"name":"Sejong University, Seoul, Republic of Korea"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2022,10,6]]},"reference":[{"key":"e_1_3_1_2_2","first-page":"1","volume-title":"Introduction to Big Multimodal Multimedia Data with Deep Analytics","author":"Wang Y.","year":"2021","unstructured":"Y. Wang, M. Fang, J. Tianyi Zhou, T. Mu, and D. Tao. 2021. Introduction to Big Multimodal Multimedia Data with Deep Analytics. 17, ed: ACM New York, NY, 2021, 1\u20133."},{"key":"e_1_3_1_3_2","doi-asserted-by":"publisher","DOI":"10.1109\/TNSE.2018.2843326"},{"key":"e_1_3_1_4_2","doi-asserted-by":"publisher","DOI":"10.1145\/3444693"},{"key":"e_1_3_1_5_2","doi-asserted-by":"publisher","DOI":"10.1145\/3449359"},{"key":"e_1_3_1_6_2","first-page":"1","volume-title":"Introduction to the Special Issue on Computational Intelligence for Biomedical Data and Imaging","author":"Tanveer M.","year":"2020","unstructured":"M. Tanveer, P. Khanna, M. Prasad, and C. Lin. 2020. Introduction to the Special Issue on Computational Intelligence for Biomedical Data and Imaging. 16, ed: ACM New York, NY, USA, 2020, 1\u20134."},{"issue":"2","key":"e_1_3_1_7_2","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/3415155","article-title":"eDiaPredict: An ensemble-based framework for diabetes prediction","volume":"17","author":"Singh A.","year":"2021","unstructured":"A. Singh, A. Dhillon, N. Kumar, M. S. Hossain, G. Muhammad, and M. Kumar. 2021. eDiaPredict: An ensemble-based framework for diabetes prediction. ACM Transactions on Multimedia Computing Communications and Applications 17, 2s (2021), 1\u201326.","journal-title":"ACM Transactions on Multimedia Computing Communications and Applications"},{"key":"e_1_3_1_8_2","doi-asserted-by":"publisher","DOI":"10.1145\/3454009"},{"key":"e_1_3_1_9_2","doi-asserted-by":"publisher","DOI":"10.1145\/3473037"},{"key":"e_1_3_1_10_2","doi-asserted-by":"publisher","DOI":"10.1145\/3418204"},{"key":"e_1_3_1_11_2","doi-asserted-by":"publisher","DOI":"10.1109\/TVT.2021.3122257"},{"key":"e_1_3_1_12_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.inffus.2021.04.016"},{"key":"e_1_3_1_13_2","doi-asserted-by":"publisher","DOI":"10.1145\/3462777"},{"key":"e_1_3_1_14_2","doi-asserted-by":"publisher","DOI":"10.1109\/MNET.201.2100069"},{"key":"e_1_3_1_15_2","unstructured":"C. Education. 2021. Data engineering preparation and labeling for AI 2019. https:\/\/www.cognilytica.com\/document\/report-data-engineering-preparation-and-labeling-for-ai-2019\/ (accessed 29\/11\/2021 2021)."},{"key":"e_1_3_1_16_2","unstructured":"I. Grand View Research. 2021. Data collection and labeling market worth $8.22 billion by 2028. https:\/\/www.grandviewresearch.com\/press-release\/global-data-collection-labeling-market (accessed 29\/11\/2021)."},{"key":"e_1_3_1_17_2","doi-asserted-by":"publisher","DOI":"10.1145\/3472291"},{"key":"e_1_3_1_18_2","doi-asserted-by":"publisher","DOI":"10.1109\/TCSVT.2016.2589879"},{"key":"e_1_3_1_19_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-69984-0_43"},{"key":"e_1_3_1_20_2","unstructured":"B. Settles. 2009. Active learning literature survey. 2009."},{"key":"e_1_3_1_21_2","doi-asserted-by":"publisher","DOI":"10.1109\/TGRS.2008.2010404"},{"key":"e_1_3_1_22_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-64148-1_13"},{"key":"e_1_3_1_23_2","unstructured":"O. Sener and S. Savarese. 2017. Active learning for convolutional neural networks: A core-set approach. arXiv preprint arXiv:1708.00489 ."},{"key":"e_1_3_1_24_2","unstructured":"S. Ebrahimi et al. 2020. Minimax active learning. arXiv preprint arXiv:2012.10467 ."},{"key":"e_1_3_1_25_2","first-page":"2016","article-title":"Budgeted stream-based active learning via adaptive submodular maximization","volume":"29","author":"Fujii K.","year":"2016","unstructured":"K. Fujii and H. Kashima. 2016. Budgeted stream-based active learning via adaptive submodular maximization. Advances in Neural Information Processing Systems 29, 2016.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_1_26_2","unstructured":"M. Ducoffe and F. Precioso. 2018. Adversarial active learning for deep networks: A margin based approach. arXiv preprint arXiv:1802.09841 ."},{"key":"e_1_3_1_27_2","first-page":"6295","volume-title":"International Conference on Machine Learning","author":"Tran T.","year":"2019","unstructured":"T. Tran, T.-T. Do, I. Reid, and G. Carneiro. 2019. Bayesian generative active deep learning. In International Conference on Machine Learning. PMLR, 6295\u20136304."},{"key":"e_1_3_1_28_2","article-title":"Self-paced learning for latent variable models","volume":"23","author":"Kumar M.","year":"2010","unstructured":"M. Kumar, B. Packer, and D. Koller. 2010. Self-paced learning for latent variable models. Advances in Neural Information Processing Systems 23, (2010).","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_1_29_2","doi-asserted-by":"publisher","DOI":"10.1145\/2505515.2505528"},{"key":"e_1_3_1_30_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.ins.2014.04.034"},{"key":"e_1_3_1_31_2","doi-asserted-by":"publisher","DOI":"10.1007\/s11390-020-9487-4"},{"key":"e_1_3_1_32_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.ins.2018.05.014"},{"key":"e_1_3_1_33_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2013.116"},{"key":"e_1_3_1_34_2","doi-asserted-by":"publisher","DOI":"10.1007\/s13748-021-00230-w"},{"key":"e_1_3_1_35_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2013.33"},{"key":"e_1_3_1_36_2","first-page":"59","volume-title":"Proceedings of the 20th International Conference on Machine Learning (ICML'03)","author":"Brinker K.","year":"2003","unstructured":"K. Brinker. 2003. Incorporating diversity in active learning with support vector machines. In Proceedings of the 20th International Conference on Machine Learning (ICML'03). 59\u201366."},{"key":"e_1_3_1_37_2","first-page":"3071","article-title":"Adversarial sampling for active learning","author":"Mayer C.","year":"2020","unstructured":"C. Mayer and R. Timofte. 2020. Adversarial sampling for active learning. In Proceedings of the IEEE\/CVF Winter Conference on Applications of Computer Vision. 3071\u20133079.","journal-title":"Proceedings of the IEEE\/CVF Winter Conference on Applications of Computer Vision"},{"key":"e_1_3_1_38_2","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-014-0781-x"},{"key":"e_1_3_1_39_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.00807"},{"key":"e_1_3_1_40_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.neucom.2020.11.019"},{"key":"e_1_3_1_41_2","doi-asserted-by":"crossref","unstructured":"J. W. Cho D.-J. Kim Y. Jung and I. S. Kweon. 2021. MCDAL: Maximum classifier discrepancy for active learning. arXiv preprint arXiv:2107.11049 .","DOI":"10.1109\/TNNLS.2022.3152786"},{"key":"e_1_3_1_42_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.90"},{"key":"e_1_3_1_43_2","volume-title":"Pattern Recognition and Neural Networks","author":"Ripley B. D.","year":"2007","unstructured":"B. D. Ripley. 2007. Pattern Recognition and Neural Networks. Cambridge University Press (2007)."},{"key":"e_1_3_1_44_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.patcog.2019.01.006"},{"key":"e_1_3_1_45_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01424-7_27"},{"key":"e_1_3_1_46_2","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-015-0816-y"},{"key":"e_1_3_1_47_2","first-page":"5998","volume-title":"Advances in Neural Information Processing Systems","author":"Vaswani A.","year":"2017","unstructured":"A. Vaswani et al. 2017. Attention is all you need. In Advances in Neural Information Processing Systems. 5998\u20136008."},{"key":"e_1_3_1_48_2","unstructured":"A. Dosovitskiy et al. 2020. An image is worth 16 \u00d7 16 words: Transformers for image recognition at scale. arXiv preprint arXiv:2010.11929 ."},{"key":"e_1_3_1_49_2","doi-asserted-by":"crossref","unstructured":"S. Khan M. Naseer M. Hayat S. W. Zamir F. S. Khan and M. Shah. 2021. Transformers in vision: A survey. arXiv preprint arXiv:2101.01169 .","DOI":"10.1145\/3505244"},{"key":"e_1_3_1_50_2","unstructured":"C. Schr\u00f6der A. Niekler and M. Potthast. Revisiting Uncertainty-based Query Strategies for Active Learning with Transformers ."},{"key":"e_1_3_1_51_2","unstructured":"A. Krizhevsky and G. Hinton. 2009. Learning multiple layers of features from tiny images. (2009)."},{"key":"e_1_3_1_52_2","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2006.79"},{"key":"e_1_3_1_53_2","unstructured":"G. Griffin A. Holub and P. Perona. 2007. Caltech-256 object category dataset. (2007)."},{"key":"e_1_3_1_54_2","unstructured":"K. Simonyan and A. Zisserman. 2014. Very deep convolutional networks for large-scale image recognition. arXiv preprint arXiv:1409.1556 ."},{"key":"e_1_3_1_55_2","volume-title":"Presented at the Proceedings of the 36th International Conference on Machine Learning, Proceedings of Machine Learning Research","author":"Tan M.","year":"2019","unstructured":"M. Tan and Q. Le. 2019. EfficientNet: Rethinking model scaling for convolutional neural networks. Presented at the Proceedings of the 36th International Conference on Machine Learning, Proceedings of Machine Learning Research (2019). [Online]. Available: https:\/\/proceedings.mlr.press\/v97\/tan19a.html."}],"container-title":["ACM Transactions on Multimedia Computing, Communications, and Applications"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3534932","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3534932","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T19:02:54Z","timestamp":1750186974000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3534932"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,6,30]]},"references-count":54,"journal-issue":{"issue":"2s","published-print":{"date-parts":[[2022,6,30]]}},"alternative-id":["10.1145\/3534932"],"URL":"https:\/\/doi.org\/10.1145\/3534932","relation":{},"ISSN":["1551-6857","1551-6865"],"issn-type":[{"type":"print","value":"1551-6857"},{"type":"electronic","value":"1551-6865"}],"subject":[],"published":{"date-parts":[[2022,6,30]]},"assertion":[{"value":"2021-12-31","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2022-04-29","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2022-10-06","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}