{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,10,10]],"date-time":"2025-10-10T01:10:26Z","timestamp":1760058626904,"version":"build-2065373602"},"reference-count":28,"publisher":"MDPI AG","issue":"4","license":[{"start":{"date-parts":[[2025,4,16]],"date-time":"2025-04-16T00:00:00Z","timestamp":1744761600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"Natural Science Foundation of Hebei Province","award":["F2024201012"],"award-info":[{"award-number":["F2024201012"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Information"],"abstract":"<jats:p>Deep learning models performing complex tasks require the support of datasets. With the advancement of virtual reality technology, the use of virtual datasets in deep learning models is becoming more and more widespread. Indoor scenes represents a significant area of interest for the application of machine vision technologies. Existing virtual indoor datasets exhibit deficiencies with regard to camera poses, resulting in problems such as occlusion, object omission, and objects having too small of a proportion of the image, and perform poorly in the training for object detection and simultaneous localization and mapping (SLAM) tasks. Aiming at the problems regarding the capacity of cameras to comprehensively capture scene objects, this study presents an enhanced algorithm based on rapidly exploring random tree star (RRT*) for the generation of camera poses in a 3D indoor scene. Meanwhile, in order to generate multimodal data for various deep learning tasks, this study designs an automatic image acquisition module under the Unity3D platform. The experimental results from running the model on several mainstream virtual indoor datasets\u2014such as 3D-FRONT and Hypersim\u2014indicate that the image sequences generated in this study show enhancements in terms of object capture rate and efficiency. Even in cluttered environments such as those in SceneNet RGB-D, the object capture rate remains stable at around 75%. Compared with the image sequences from the original datasets, those generated in this study achieve improvements in the object detection and SLAM tasks, with increases of up to approximately 30% in mAP for the YOLOv10 object detection task and up to approximately 10% in SR for the ORB-SLAM algorithm.<\/jats:p>","DOI":"10.3390\/info16040315","type":"journal-article","created":{"date-parts":[[2025,4,16]],"date-time":"2025-04-16T08:09:17Z","timestamp":1744790957000},"page":"315","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":0,"title":["Camera Pose Generation Based on Unity3D"],"prefix":"10.3390","volume":"16","author":[{"ORCID":"https:\/\/orcid.org\/0009-0003-2599-6908","authenticated-orcid":false,"given":"Hao","family":"Luo","sequence":"first","affiliation":[{"name":"School of Cyber Security and Computer, Hebei University, Baoding 071002, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-2070-465X","authenticated-orcid":false,"given":"Wenjie","family":"Luo","sequence":"additional","affiliation":[{"name":"School of Cyber Security and Computer, Hebei University, Baoding 071002, China"},{"name":"Machine Vision Engineering Research Center, Hebei University, Baoding 071002, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Wenzhu","family":"Yang","sequence":"additional","affiliation":[{"name":"School of Cyber Security and Computer, Hebei University, Baoding 071002, China"},{"name":"Machine Vision Engineering Research Center, Hebei University, Baoding 071002, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1968","published-online":{"date-parts":[[2025,4,16]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","unstructured":"Macario Barros, A., Michel, M., Moline, Y., Corre, G., and Carrel, F. (2022). A comprehensive survey of visual slam algorithms. Robotics, 11.","DOI":"10.3390\/robotics11010024"},{"key":"ref_2","unstructured":"Ramakrishnan, S.K., Gokaslan, A., Wijmans, E., Maksymets, O., Clegg, A., Turner, J., Undersander, E., Galuba, W., Westbury, A., and Chang, A.X. (2021). Habitat-matterport 3D dataset (HM3D): 1000 large-scale 3D environments for embodied AI. arXiv."},{"key":"ref_3","unstructured":"McCormac, J., Handa, A., Leutenegger, S., and Davison, A.J. (2016). Scenenet rgb-d: 5m photorealistic images of synthetic indoor trajectories with ground truth. arXiv."},{"key":"ref_4","doi-asserted-by":"crossref","unstructured":"Fu, H., Cai, B., Gao, L., Zhang, L.X., Wang, J., Li, C., Zeng, Q., Sun, C., Jia, R., and Zhao, B. (2021, January 11\u201317). 3d-front: 3d furnished rooms with layouts and semantics. Proceedings of the IEEE\/CVF International Conference on Computer Vision, Montreal, BC, Canada.","DOI":"10.1109\/ICCV48922.2021.01075"},{"key":"ref_5","doi-asserted-by":"crossref","unstructured":"Wang, W., Zhu, D., Wang, X., Hu, Y., Qiu, Y., Wang, C., Hu, Y., Kapoor, A., and Scherer, S. (2020, January 24\u201328). Tartanair: A dataset to push the limits of visual slam. Proceedings of the 2020 IEEE\/RSJ International Conference on Intelligent Robots and Systems (IROS), Las Vegas, NV, USA.","DOI":"10.1109\/IROS45743.2020.9341801"},{"key":"ref_6","doi-asserted-by":"crossref","unstructured":"Tasora, A., Serban, R., Mazhar, H., Pazouki, A., Melanz, D., Fleischmann, J., Taylor, M., Sugiyama, H., and Negrut, D. (2015, January 25\u201328). Chrono: An open source multi-physics dynamics engine. Proceedings of the High Performance Computing in Science and Engineering: Second International Conference, HPCSE 2015, Sol\u00e1\u0148, Czech Republic.","DOI":"10.1007\/978-3-319-40361-8_2"},{"key":"ref_7","doi-asserted-by":"crossref","first-page":"12562","DOI":"10.1109\/ACCESS.2024.3354709","article-title":"A comprehensive survey on Delaunay triangulation: Applications, algorithms, and implementations over CPUs, GPUs, and FPGAs","volume":"12","author":"Elshakhs","year":"2024","journal-title":"IEEE Access"},{"key":"ref_8","doi-asserted-by":"crossref","unstructured":"Noreen, I., Khan, A., and Habib, Z. (2016). Optimal path planning using RRT* based approaches: A survey and future directions. Int. J. Adv. Comput. Sci. Appl., 7.","DOI":"10.14569\/IJACSA.2016.071114"},{"key":"ref_9","doi-asserted-by":"crossref","first-page":"1231","DOI":"10.1177\/0278364913491297","article-title":"Vision meets robotics: The kitti dataset","volume":"32","author":"Geiger","year":"2013","journal-title":"Int. J. Robot. Res."},{"key":"ref_10","unstructured":"Du, W., and Beltrame, G. (2024). LiDAR-based Real-Time Object Detection and Tracking in Dynamic Environments. arXiv."},{"key":"ref_11","doi-asserted-by":"crossref","unstructured":"Srivastava, H.M. (2023). An introductory overview of Bessel polynomials, the generalized Bessel polynomials and the q-Bessel polynomials. Symmetry, 15.","DOI":"10.3390\/sym15040822"},{"key":"ref_12","doi-asserted-by":"crossref","unstructured":"Silberman, N., Hoiem, D., Kohli, P., and Fergus, R. (2012, January 7\u201313). Indoor segmentation and support inference from rgbd images. Proceedings of the Computer Vision\u2013ECCV 2012: 12th European Conference on Computer Vision, Florence, Italy.","DOI":"10.1007\/978-3-642-33715-4_54"},{"key":"ref_13","unstructured":"Li, W., Saeedi, S., McCormac, J., Clark, R., Tzoumanikas, D., Ye, Q., Huang, Y., Tang, R., and Leutenegger, S. (2018). Interiornet: Mega-scale multi-sensor photo-realistic indoor scenes dataset. arXiv."},{"key":"ref_14","doi-asserted-by":"crossref","unstructured":"Roberts, M., Ramapuram, J., Ranjan, A., Kumar, A., Bautista, M.A., Paczan, N., Webb, R., and Susskind, J.M. (2021, January 11\u201317). Hypersim: A photorealistic synthetic dataset for holistic indoor scene understanding. Proceedings of the IEEE\/CVF international Conference on Computer Vision, Montreal, BC, Canada.","DOI":"10.1109\/ICCV48922.2021.01073"},{"key":"ref_15","doi-asserted-by":"crossref","unstructured":"Gharehbagh, A.K., Judeh, R., Ng, J., von Reventlow, C., and R\u00f6hrbein, F. (2021, January 11\u201314). Real-time 3D Semantic Mapping based on Keyframes and Octomap for Autonomous Cobot. Proceedings of the 2021 9th International Conference on Control, Mechatronics and Automation (ICCMA), Luxembourg.","DOI":"10.1109\/ICCMA54375.2021.9646203"},{"key":"ref_16","doi-asserted-by":"crossref","unstructured":"Pivo\u0148ka, T., and P\u0159eu\u010dil, L. (2020, January 21). Stereo camera simulation in blender. Proceedings of the Modelling and Simulation for Autonomous Systems: 7th International Conference, MESAS 2020, Prague, Czech Republic.","DOI":"10.1007\/978-3-030-70740-8_13"},{"key":"ref_17","doi-asserted-by":"crossref","first-page":"173","DOI":"10.1111\/cgf.14077","article-title":"Poisson surface reconstruction with envelope constraints","volume":"Volume 39","author":"Kazhdan","year":"2020","journal-title":"Computer Graphics Forum"},{"key":"ref_18","unstructured":"Barrera, T., Hast, A., and Bengtsson, E. (2004, January 24\u201325). Incremental spherical linear interpolation. Proceedings of the The Annual SIGRAD Conference. Special Theme-Environmental Visualization, G\u00e4vle, Sweden."},{"key":"ref_19","doi-asserted-by":"crossref","first-page":"167","DOI":"10.54097\/fm9v2e71","article-title":"Game Design Based on High-Definition Render Pipeline in Unity","volume":"93","author":"Peng","year":"2024","journal-title":"Highlights Sci. Eng. Technol."},{"key":"ref_20","doi-asserted-by":"crossref","unstructured":"Zheng, Z., Wang, P., Liu, W., Li, J., Ye, R., and Ren, D. (2020, January 7\u201312). Distance-IoU loss: Faster and better learning for bounding box regression. Proceedings of the AAAI Conference on Artificial Intelligence, New York, NY, USA.","DOI":"10.1609\/aaai.v34i07.6999"},{"key":"ref_21","unstructured":"Chang, A.X., Funkhouser, T., Guibas, L., Hanrahan, P., Huang, Q., Li, Z., Savarese, S., Savva, M., Song, S., and Su, H. (2015). Shapenet: An information-rich 3d model repository. arXiv."},{"key":"ref_22","unstructured":"(2025, February 26). TurboSquid. Available online: http:\/\/www.turbosquid.com\/."},{"key":"ref_23","unstructured":"Wang, A., Chen, H., Liu, L., Chen, K., Lin, Z., Han, J., and Ding, G. (2024). Yolov10: Real-time end-to-end object detection. arXiv."},{"key":"ref_24","doi-asserted-by":"crossref","unstructured":"Zhao, Y., Lv, W., Xu, S., Wei, J., Wang, G., Dang, Q., Liu, Y., and Chen, J. (2024, January 17\u201318). Detrs beat yolos on real-time object detection. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Seattle, WA, USA.","DOI":"10.1109\/CVPR52733.2024.01605"},{"key":"ref_25","first-page":"821","article-title":"A review of non-maximum suppression algorithms for deep learning target detection","volume":"Volume 11763","author":"Gong","year":"2021","journal-title":"Proceedings of the Seventh Symposium on Novel Photoelectronic Detection Technology and Applications"},{"key":"ref_26","doi-asserted-by":"crossref","first-page":"1147","DOI":"10.1109\/TRO.2015.2463671","article-title":"ORB-SLAM: A versatile and accurate monocular SLAM system","volume":"31","author":"Montiel","year":"2015","journal-title":"IEEE Trans. Robot."},{"key":"ref_27","doi-asserted-by":"crossref","unstructured":"Engel, J., Sch\u00f6ps, T., and Cremers, D. (2014, January 6\u201312). LSD-SLAM: Large-scale direct monocular SLAM. Proceedings of the European Conference on Computer Vision, Zurich, Switzerland.","DOI":"10.1007\/978-3-319-10605-2_54"},{"key":"ref_28","doi-asserted-by":"crossref","first-page":"611","DOI":"10.1109\/TPAMI.2017.2658577","article-title":"Direct sparse odometry","volume":"40","author":"Engel","year":"2017","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."}],"container-title":["Information"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/2078-2489\/16\/4\/315\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,9]],"date-time":"2025-10-09T17:15:40Z","timestamp":1760030140000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/2078-2489\/16\/4\/315"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,4,16]]},"references-count":28,"journal-issue":{"issue":"4","published-online":{"date-parts":[[2025,4]]}},"alternative-id":["info16040315"],"URL":"https:\/\/doi.org\/10.3390\/info16040315","relation":{},"ISSN":["2078-2489"],"issn-type":[{"type":"electronic","value":"2078-2489"}],"subject":[],"published":{"date-parts":[[2025,4,16]]}}}