{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,10,12]],"date-time":"2025-10-12T02:15:03Z","timestamp":1760235303203,"version":"build-2065373602"},"reference-count":28,"publisher":"MDPI AG","issue":"8","license":[{"start":{"date-parts":[[2021,8,14]],"date-time":"2021-08-14T00:00:00Z","timestamp":1628899200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["J. Imaging"],"abstract":"<jats:p>Collecting real-world data for the training of neural networks is enormously time-consuming and expensive. As such, the concept of virtualizing the domain and creating synthetic data has been analyzed in many instances. This virtualization offers many possibilities of changing the domain, and with that, enabling the relatively fast creation of data. It also offers the chance to enhance necessary augmentations with additional semantic information when compared with conventional augmentation methods. This raises the question of whether such semantic changes, which can be seen as augmentations of the virtual domain, contribute to better results for neural networks, when trained with data augmented this way. In this paper, a virtual dataset is presented, including semantic augmentations and automatically generated annotations, as well as a comparison between semantic and conventional augmentation for image data. It is determined that the results differ only marginally for neural network models trained with the two augmentation approaches.<\/jats:p>","DOI":"10.3390\/jimaging7080146","type":"journal-article","created":{"date-parts":[[2021,8,15]],"date-time":"2021-08-15T21:43:55Z","timestamp":1629063835000},"page":"146","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":1,"title":["Investigating Semantic Augmentation in Virtual Environments for Image Segmentation Using Convolutional Neural Networks"],"prefix":"10.3390","volume":"7","author":[{"given":"Joshua","family":"Ganter","sequence":"first","affiliation":[{"name":"Faculty of Digital Media, Furtwangen University, 78120 Furtwangen, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-2443-5549","authenticated-orcid":false,"given":"Simon","family":"L\u00f6ffler","sequence":"additional","affiliation":[{"name":"Faculty of Digital Media, Furtwangen University, 78120 Furtwangen, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Ron","family":"Metzger","sequence":"additional","affiliation":[{"name":"Faculty of Digital Media, Furtwangen University, 78120 Furtwangen, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Katharina","family":"U\u00dfling","sequence":"additional","affiliation":[{"name":"Faculty of Digital Media, Furtwangen University, 78120 Furtwangen, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Christoph","family":"M\u00fcller","sequence":"additional","affiliation":[{"name":"Faculty of Digital Media, Furtwangen University, 78120 Furtwangen, Germany"},{"name":"Fraunhofer-Institute for Physical Measurement Techniques IPM, 79110 Freiburg, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1968","published-online":{"date-parts":[[2021,8,14]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","first-page":"1089","DOI":"10.1007\/s10462-018-9641-3","article-title":"Recent progress in semantic image segmentation","volume":"52","author":"Liu","year":"2019","journal-title":"Artif. Intell. Rev."},{"doi-asserted-by":"crossref","unstructured":"Valada, A., Vertens, J., Dhall, A., and Burgard, W. (June, January 29). AdapNet: Adaptive Semantic Segmentation in Adverse Environmental Conditions. Proceedings of the 2017 IEEE International Conference on Robotics and Automation (ICRA), Singapore.","key":"ref_2","DOI":"10.1109\/ICRA.2017.7989540"},{"unstructured":"Alexander, B.J., Kentaro, W., Jon, C., Satoshi, T., Jake, G., and Christoph, R. (2020, December 10). Imgaug. Available online: https:\/\/github.com\/aleju\/imgaug.","key":"ref_3"},{"doi-asserted-by":"crossref","unstructured":"Liu, S., Zhang, J., Chen, Y., Liu, Y., Qin, Z., and Wan, T. (2019, January 12\u201317). Pixel Level Data Augmentation for Semantic Image Segmentation Using Generative Adversarial Networks. Proceedings of the ICASSP 2019\u20142019 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), Brighton, UK.","key":"ref_4","DOI":"10.1109\/ICASSP.2019.8683590"},{"key":"ref_5","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1186\/s40537-019-0197-0","article-title":"A survey on Image Data Augmentation for Deep Learning","volume":"6","author":"Shorten","year":"2019","journal-title":"J. Big Data"},{"unstructured":"Yuan, Y., Huang, L., Guo, J., Zhang, C., Chen, X., and Wang, J. (2018). OCNet: Object Context Network for Scene Parsing. arXiv.","key":"ref_6"},{"doi-asserted-by":"crossref","unstructured":"Ding, X., Guo, Y., Ding, G., and Han, J. (2019). ACNet: Strengthening the Kernel Skeletons for Powerful CNN via Asymmetric Convolution Blocks. arXiv.","key":"ref_7","DOI":"10.1109\/ICCV.2019.00200"},{"doi-asserted-by":"crossref","unstructured":"Huang, Z., Wang, X., Huang, L., Huang, C., Wei, Y., and Liu, W. (2019). CCNet: Criss-Cross Attention for Semantic Segmentation. arXiv.","key":"ref_8","DOI":"10.1109\/ICCV.2019.00069"},{"doi-asserted-by":"crossref","unstructured":"Zheng, S., Lu, J., Zhao, H., Zhu, X., Luo, Z., Wang, Y., Fu, Y., Feng, J., Xiang, T., and Torr, P.H. (2021). Rethinking Semantic Segmentation from a Sequence-to-Sequence Perspective with Transformers. arXiv.","key":"ref_9","DOI":"10.1109\/CVPR46437.2021.00681"},{"unstructured":"Xie, E., Wang, W., Yu, Z., An kumar, A., Alvarez, J.M., and Luo, P. (2021). SegFormer: Simple and Efficient Design for Semantic Segmentation with Transformers. arXiv.","key":"ref_10"},{"unstructured":"Cheng, B., Schwing, A.G., and Kirillov, A. (2021). Per-Pixel Classification is Not All You Need for Semantic Segmentation. arXiv.","key":"ref_11"},{"doi-asserted-by":"crossref","unstructured":"Zendel, O., Murschitz, M., Zeilinger, M., Steininger, D., Abbasi, S., and Beleznai, C. (2019, January 16\u201317). RailSem19: A Dataset for Semantic Rail Scene Understanding. Proceedings of the 2019 IEEE\/CVF Conference on Computer, Long Beach, CA, USA.","key":"ref_12","DOI":"10.1109\/CVPRW.2019.00161"},{"unstructured":"Harb, J., R\u00e9b\u00e9na, N., Chosidow, R., Roblin, G., Potarusov, R., and Hajri, H. (2020). FRSign: A Large-Scale Traffic Light Dataset for Autonomous Trains. arXiv.","key":"ref_13"},{"doi-asserted-by":"crossref","unstructured":"Gaidon, A., Wang, Q., Cabon, Y., and Vig, E. (2016). Virtual Worlds as Proxy for Multi-Object Tracking Analysis. arXiv.","key":"ref_14","DOI":"10.1109\/CVPR.2016.470"},{"unstructured":"Cabon, Y., Murray, N., and Humenberger, M. (2020). Virtual KITTI 2. arXiv.","key":"ref_15"},{"key":"ref_16","doi-asserted-by":"crossref","first-page":"526","DOI":"10.1197\/jamia.M2051","article-title":"Enhancing text categorization with semantic-enriched representation and training data augmentation","volume":"13","author":"Lu","year":"2006","journal-title":"J. Am. Med. Inform. Assoc."},{"key":"ref_17","doi-asserted-by":"crossref","first-page":"7176","DOI":"10.1609\/aaai.v33i01.33017176","article-title":"A Task in a Suit and a Tie: Paraphrase Generation with Semantic Augmentation","volume":"33","author":"Wang","year":"2019","journal-title":"Proc. AAAI"},{"unstructured":"Wang, Y., Pan, X., Song, S., Zhang, H., Wu, C., and Huang, G. (2019). Implicit Semantic Data Augmentation for Deep Networks; 33rd Conference on Neural Information Processing Systems (NeurIPS 2019). arXiv.","key":"ref_18"},{"unstructured":"Wood, E. (2021, August 05). Synthetic Data with Digital Humans; Microsoft. Available online: https:\/\/www.microsoft.com\/en-us\/research\/uploads\/prod\/2019\/09\/2019-10-01-Synthetic-Data-with-Digital-Humans.pdf.","key":"ref_19"},{"doi-asserted-by":"crossref","unstructured":"Zhang, R., Isola, P., Efros, A.A., Shechtman, E., and Wang, O. (2018). The Unreasonable Effectiveness of Deep Features as a Perceptual Metric. arXiv.","key":"ref_20","DOI":"10.1109\/CVPR.2018.00068"},{"doi-asserted-by":"crossref","unstructured":"He, K., Zhang, X., Ren, S., and Sun, J. (2015). Deep Residual Learning for Image Recognition. arXiv.","key":"ref_21","DOI":"10.1109\/CVPR.2016.90"},{"doi-asserted-by":"crossref","unstructured":"Everingham, M., Van Gool, L., Williams, C.K.I., Winn, J., and Zisserman, A. (2010). The pascal visual object classes (VOC) challenge. Int. J. Comput. Vis.","key":"ref_22","DOI":"10.1007\/s11263-009-0275-4"},{"unstructured":"Blum, H., Sarlin, P.-E., Nieto, J., Siegwart, R., and Cadena, C. (2019). The Fishyscapes Benchmark: Measuring Blind Spots in Semantic Segmentation. arXiv.","key":"ref_23"},{"doi-asserted-by":"crossref","unstructured":"Zhang, J., Yang, K., and Stiefelhagen, R. (2020). ISSAFE: Improving Semantic Segmentation in Accidents by Fusing Event-based Data. arXiv.","key":"ref_24","DOI":"10.1109\/IROS51168.2021.9636109"},{"doi-asserted-by":"crossref","unstructured":"Zendel, O., Honauer, K., Murschitz, M., Steininger, D., and Dominguez, G.F. (2018, January 8\u201314). WildDash\u2014Creating Hazard-Aware Benchmarks; Springer International Publishing. Proceedings of the European Conference on Computer Vision (ECCV), Munich, Germany.","key":"ref_25","DOI":"10.1007\/978-3-030-01231-1_25"},{"doi-asserted-by":"crossref","unstructured":"Sakaridis, C., Dai, D., and van Gool, L. (2021). ACDC: The Adverse Conditions Dataset with Correspondences for Semantic Driving Scene Understanding. arXiv.","key":"ref_26","DOI":"10.1109\/ICCV48922.2021.01059"},{"unstructured":"(2021, August 05). AIT Austrian Institute of Technology (2021): WildDash 2 Benchmark, RailSem19: A Dataset for Semantic Rail Scene Understanding. Available online: https:\/\/wilddash.cc\/railsem19.","key":"ref_27"},{"unstructured":"L\u00f6ffler, S., Metzger, R., U\u00dfling, K., and Ganter, J. (2021). SynTra train ride dataset. figshare.","key":"ref_28"}],"container-title":["Journal of Imaging"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/2313-433X\/7\/8\/146\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,11]],"date-time":"2025-10-11T06:46:14Z","timestamp":1760165174000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/2313-433X\/7\/8\/146"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2021,8,14]]},"references-count":28,"journal-issue":{"issue":"8","published-online":{"date-parts":[[2021,8]]}},"alternative-id":["jimaging7080146"],"URL":"https:\/\/doi.org\/10.3390\/jimaging7080146","relation":{},"ISSN":["2313-433X"],"issn-type":[{"type":"electronic","value":"2313-433X"}],"subject":[],"published":{"date-parts":[[2021,8,14]]}}}