{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,4]],"date-time":"2026-04-04T18:13:06Z","timestamp":1775326386364,"version":"3.50.1"},"reference-count":33,"publisher":"MDPI AG","issue":"12","license":[{"start":{"date-parts":[[2018,12,8]],"date-time":"2018-12-08T00:00:00Z","timestamp":1544227200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"National Key R&amp;G Program of China","award":["2018YFB1004600"],"award-info":[{"award-number":["2018YFB1004600"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Symmetry"],"abstract":"<jats:p>Augmented Reality (AR) is crucial for immersive Human\u2013Computer Interaction (HCI) and the vision of Artificial Intelligence (AI). Labeled data drives object recognition in AR. However, manually annotating data is expensive, labor-intensive, and data distribution asymmetry. Scantily labeled data limits the application of AR. Aiming at solving the problem of insufficient and asymmetry training data in AR object recognition, an automated vision data synthesis method, i.e., background augmentation generative adversarial networks (BAGANs), is proposed in this paper based on 3D modeling and the Generative Adversarial Network (GAN) algorithm. Our approach has been validated to have better performance than other methods through image recognition tasks with respect to the natural image database ObjectNet3D. This study can shorten the algorithm development time of AR and expand its application scope, which is of great significance for immersive interactive systems.<\/jats:p>","DOI":"10.3390\/sym10120734","type":"journal-article","created":{"date-parts":[[2018,12,10]],"date-time":"2018-12-10T03:36:41Z","timestamp":1544413001000},"page":"734","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":30,"title":["Background Augmentation Generative Adversarial Networks (BAGANs): Effective Data Generation Based on GAN-Augmented 3D Synthesizing"],"prefix":"10.3390","volume":"10","author":[{"ORCID":"https:\/\/orcid.org\/0000-0001-9243-8322","authenticated-orcid":false,"given":"Yan","family":"Ma","sequence":"first","affiliation":[{"name":"School of Mechanical Electronic &amp; Information Engineering, China University of Mining &amp; Technology, Beijing 100083, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-8413-123X","authenticated-orcid":false,"given":"Kang","family":"Liu","sequence":"additional","affiliation":[{"name":"School of Mechanical Electronic &amp; Information Engineering, China University of Mining &amp; Technology, Beijing 100083, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Zhibin","family":"Guan","sequence":"additional","affiliation":[{"name":"School of Mechanical Electronic &amp; Information Engineering, China University of Mining &amp; Technology, Beijing 100083, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xinkai","family":"Xu","sequence":"additional","affiliation":[{"name":"School of Mechanical Electronic &amp; Information Engineering, China University of Mining &amp; Technology, Beijing 100083, China"},{"name":"Demonstration Center of Experimental Teaching in Comprehensive Engineering, Beijing Union University, Beijing 100101, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xu","family":"Qian","sequence":"additional","affiliation":[{"name":"School of Mechanical Electronic &amp; Information Engineering, China University of Mining &amp; Technology, Beijing 100083, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Hong","family":"Bao","sequence":"additional","affiliation":[{"name":"Beijing Key Laboratory of Information Service Engineering, Beijing Union University, Beijing 100101, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1968","published-online":{"date-parts":[[2018,12,8]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","unstructured":"Richer, R., Maiwald, T., Pasluosta, C., Hensel, B., and Eskofier, B.M. (2015, January 9\u201312). Novel human computer interaction principles for cardiac feedback using google glass and Android wear. Proceedings of the 2015 IEEE 12th International Conference on Wearable and Implantable Body Sensor Networks (BSN), Cambridge, MA, USA.","DOI":"10.1109\/BSN.2015.7299363"},{"key":"ref_2","first-page":"10","article-title":"Considering privacy issues in the context of Google glass","volume":"56","author":"Hong","year":"2013","journal-title":"Commun. ACM"},{"key":"ref_3","doi-asserted-by":"crossref","unstructured":"Evans, G., Miller, J., Pena, M.I., Macallister, A., and Winer, E.H. (2017). Evaluating the Microsoft HoloLens through an augmented reality assembly application. Proc. SPIE, 10197.","DOI":"10.1117\/12.2262626"},{"key":"ref_4","doi-asserted-by":"crossref","first-page":"146","DOI":"10.1016\/j.inffus.2017.10.006","article-title":"A survey on deep learning for big data","volume":"42","author":"Zhang","year":"2018","journal-title":"Inf. Fusion"},{"key":"ref_5","doi-asserted-by":"crossref","first-page":"101","DOI":"10.1016\/j.patrec.2018.04.010","article-title":"Hybrid deep neural networks for face emotion recognition","volume":"115","author":"Jain","year":"2018","journal-title":"Pattern Recognit. Lett."},{"key":"ref_6","doi-asserted-by":"crossref","first-page":"1031","DOI":"10.1007\/s11045-016-0429-9","article-title":"Fake modern Chinese painting identification based on spectral\u2013spatial feature fusion on hyperspectral image","volume":"27","author":"Wang","year":"2016","journal-title":"Multidimens. Syst. Signal Process."},{"key":"ref_7","doi-asserted-by":"crossref","unstructured":"Sun, M., Zhang, D., Ren, J., Wang, Z., and Jin, J.S. (2015, January 27\u201330). Brushstroke based sparse hybrid convolutional neural networks for author classification of Chinese ink-wash paintings. Proceedings of the 2015 IEEE International Conference on Image Processing (ICIP), Quebec City, QC, Canada.","DOI":"10.1109\/ICIP.2015.7350874"},{"key":"ref_8","doi-asserted-by":"crossref","unstructured":"He, K., Gkioxari, G., Doll\u00e1r, P., and Girshick, R. (2017, January 22\u201329). Mask r-cnn. Proceedings of the 2017 IEEE International Conference on Computer Vision (ICCV), Venice, Italy.","DOI":"10.1109\/ICCV.2017.322"},{"key":"ref_9","unstructured":"Krizhevsky, A., Sutskever, I., and Hinton, G.E. (2012, January 3\u20136). ImageNet classification with deep convolutional neural networks. Proceedings of the International Conference on Neural Information Processing Systems, Lake Tahoe, NV, USA."},{"key":"ref_10","doi-asserted-by":"crossref","unstructured":"Szegedy, C., Liu, W., Jia, Y., Sermanet, P., Reed, S.E., Anguelov, D., Erhan, D., Vanhoucke, V., and Rabinovich, A. (2015, January 7\u201312). Going deeper with convolutions. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Boston, MA, USA.","DOI":"10.1109\/CVPR.2015.7298594"},{"key":"ref_11","doi-asserted-by":"crossref","first-page":"611","DOI":"10.1111\/1467-9868.00196","article-title":"Probabilistic Principal Component Analysis","volume":"61","author":"Tipping","year":"1999","journal-title":"J. R. Stat. Soc. Ser. B Stat. Methodol."},{"key":"ref_12","unstructured":"Goodfellow, I.J., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., Courville, A., and Bengio, Y. (2014, January 8\u201313). Generative adversarial nets. Proceedings of the 27th International Conference on Neural Information Processing Systems, Montreal, QC, Canada."},{"key":"ref_13","unstructured":"Odena, A., Olah, C., and Shlens, J. (2017, January 6\u201311). Conditional Image Synthesis With Auxiliary Classifier GANs. Proceedings of the International Conference on Machine Learning, Sydney, Australia."},{"key":"ref_14","unstructured":"Radford, A., Metz, L., and Chintala, S. (2016, January 2\u20134). Unsupervised Representation Learning with Deep Convolutional Generative Adversarial Networks. Proceedings of the International Conference on Learning Representations, San Juan, Puerto Rico."},{"key":"ref_15","first-page":"5","article-title":"Learning from Simulated and Unsupervised Images through Adversarial Training","volume":"2","author":"Shrivastava","year":"2017","journal-title":"CVPR"},{"key":"ref_16","doi-asserted-by":"crossref","unstructured":"Zongker, D.E., Werner, D.M., Curless, B., and Salesin, D.H. (1999). Environment Matting and Compositing. Proceedings of the 26th Annual Conference on Computer Graphics and Interactive Techniques, ACM Press\/Addison-Wesley Publishing Co.","DOI":"10.1145\/311535.311558"},{"key":"ref_17","doi-asserted-by":"crossref","first-page":"253","DOI":"10.1145\/964965.808606","article-title":"Compositing Digital Images","volume":"18","author":"Porter","year":"1984","journal-title":"SIGGRAPH Comput. Graph."},{"key":"ref_18","doi-asserted-by":"crossref","unstructured":"Zhu, J.Y., Park, T., Isola, P., and Efros, A.A. (2017, January 22\u201329). Unpaired Image-to-Image Translation using Cycle-Consistent Adversarial Networks. Proceedings of the 2017 IEEE International Conference on Computer Vision (ICCV), Venice, Italy.","DOI":"10.1109\/ICCV.2017.244"},{"key":"ref_19","unstructured":"Arjovsky, M., and Bottou, L. (arXiv, 2017). Towards principled methods for training generative adversarial networks, arXiv."},{"key":"ref_20","unstructured":"Arjovsky, M., Chintala, S., and Bottou, L. (arXiv, 2017). Wasserstein gan, arXiv."},{"key":"ref_21","unstructured":"Gulrajani, I., Ahmed, F., Arjovsky, M., Dumoulin, V., and Courville, A.C. (2017, January 4\u20139). Improved training of wasserstein gans. Proceedings of the Advances in Neural Information Processing Systems, Long Beach, CA, USA."},{"key":"ref_22","unstructured":"Mirza, M., and Osindero, S. (arXiv, 2014). Conditional Generative Adversarial Nets, arXiv."},{"key":"ref_23","unstructured":"Odena, A. (arXiv, 2016). Semi-Supervised Learning with Generative Adversarial Networks, arXiv."},{"key":"ref_24","unstructured":"Chen, X., Duan, Y., Houthooft, R., Schulman, J., Sutskever, I., and Abbeel, P. (2016, January 5\u201310). InfoGAN: Interpretable Representation Learning by Information Maximizing Generative Adversarial Nets. Proceedings of the 30th International Conference on Neural Information Processing Systems, Barcelona, Spain."},{"key":"ref_25","doi-asserted-by":"crossref","unstructured":"Xiang, Y., Kim, W., Chen, W., Ji, J., Choy, C., Su, H., Mottaghi, R., Guibas, L., and Savarese, S. (2016). ObjectNet3D: A Large Scale Database for 3D Object Recognition. European Conference Computer Vision (ECCV), Springer.","DOI":"10.1007\/978-3-319-46484-8_10"},{"key":"ref_26","unstructured":"Chang, A.X., Funkhouser, T., Guibas, L., Hanrahan, P., Huang, Q., Li, Z., Savarese, S., Savva, M., Song, S., and Su, H. (arXiv, 2015). ShapeNet: An Information-Rich 3D Model Repository, arXiv."},{"key":"ref_27","unstructured":"Zhao, D., Zheng, J., and Ren, J. (September, January 31). Effective Removal of Artifacts from Views Synthesized using Depth Image Based Rendering. Proceedings of the International Conference on Distributed Multimedia Systems, Vancouver, BC, USA."},{"key":"ref_28","doi-asserted-by":"crossref","unstructured":"Smith, A.R., and Blinn, J.F. (1996). Blue Screen Matting. Proceedings of the 23rd Annual Conference on Computer Graphics and Interactive Techniques, ACM.","DOI":"10.1145\/237170.237263"},{"key":"ref_29","unstructured":"Simonyan, K., and Zisserman, A. (2015, January 7\u20139). Very Deep Convolutional Networks for Large-Scale Image Recognition. Proceedings of the International Conference on Learning Representations, San Diego, CA, USA."},{"key":"ref_30","doi-asserted-by":"crossref","unstructured":"Xiao, J., Hays, J., Ehinger, K.A., Oliva, A., and Torralba, A. (2010, January 13\u20138). SUN database: Large-scale scene recognition from abbey to zoo. Proceedings of the 2010 IEEE Computer Society Conference on Computer Vision and Pattern Recognition, San Francisco, CA, USA.","DOI":"10.1109\/CVPR.2010.5539970"},{"key":"ref_31","doi-asserted-by":"crossref","first-page":"68","DOI":"10.1016\/j.neucom.2018.01.076","article-title":"A Deep-Learning Based Feature Hybrid Framework for Spatiotemporal Saliency Detection inside Videos","volume":"287","author":"Wang","year":"2018","journal-title":"Neurocomputing"},{"key":"ref_32","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1016\/j.neucom.2015.11.044","article-title":"Novel segmented stacked autoencoder for effective dimensionality reduction and feature extraction in hyperspectral imaging","volume":"185","author":"Zabalza","year":"2016","journal-title":"Neurocomputing"},{"key":"ref_33","doi-asserted-by":"crossref","first-page":"1309","DOI":"10.1109\/TCSVT.2014.2381471","article-title":"Background Prior-Based Salient Object Detection via Deep Reconstruction Residual","volume":"25","author":"Han","year":"2015","journal-title":"IEEE Trans. Circuits Syst. Video Technol."}],"container-title":["Symmetry"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/2073-8994\/10\/12\/734\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,11]],"date-time":"2025-10-11T15:32:11Z","timestamp":1760196731000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/2073-8994\/10\/12\/734"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2018,12,8]]},"references-count":33,"journal-issue":{"issue":"12","published-online":{"date-parts":[[2018,12]]}},"alternative-id":["sym10120734"],"URL":"https:\/\/doi.org\/10.3390\/sym10120734","relation":{"has-preprint":[{"id-type":"doi","id":"10.20944\/preprints201811.0252.v1","asserted-by":"object"}]},"ISSN":["2073-8994"],"issn-type":[{"value":"2073-8994","type":"electronic"}],"subject":[],"published":{"date-parts":[[2018,12,8]]}}}