{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,4]],"date-time":"2026-06-04T05:03:46Z","timestamp":1780549426430,"version":"3.54.1"},"reference-count":37,"publisher":"MDPI AG","issue":"6","license":[{"start":{"date-parts":[[2026,5,28]],"date-time":"2026-05-28T00:00:00Z","timestamp":1779926400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"Key Research and Development Projects of Sichuan Provincial Science and Technology Program","award":["2024YFCY0029"],"award-info":[{"award-number":["2024YFCY0029"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["J. Imaging"],"abstract":"<jats:p>Regular geometric targets under microscopic scenes, such as microspheres, micropores, and microtubes, are characterized by small scales, low contrast, and degraded boundaries. Masks generated by general segmentation methods often fail to directly support high-precision geometric parameter measurement. This paper proposes a mask optimization method for the high-precision extraction of regular geometric features in microscopic scenes. We establish a mask optimization framework that integrates initial mask generation with geometric consistency refinement. Mask initialization is first performed through segmentation and adaptive super-resolution (SR) under low annotation constraints. Subsequently, an iterative optimization strategy that fuses multi-dimensional pixel features with regular geometric priors is designed. By incorporating geometric features extracted from the current mask while maintaining stable pixel-level observations, the mask is progressively corrected until convergence to generate target masks with continuous boundaries that satisfy stringent geometric constraints. Our experimental results on a sphere\u2013tube assembly dataset demonstrate that the proposed method achieves lower geometric errors on successfully fitted samples and significantly improves the fitting success rate. Ablation studies further confirm the critical roles of dynamic SR and iterative mask optimization in enhancing overall precision and stability. These findings suggest that for microscopic regular geometric measurement tasks, integrating geometric-consistency constraints into mask optimization effectively improves both the accuracy and robustness of geometric feature extraction.<\/jats:p>","DOI":"10.3390\/jimaging12060238","type":"journal-article","created":{"date-parts":[[2026,5,29]],"date-time":"2026-05-29T15:57:18Z","timestamp":1780070238000},"page":"238","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":0,"title":["Mask Optimization for High-Precision Extraction of Geometric Features in Microscopic Scenes"],"prefix":"10.3390","volume":"12","author":[{"given":"Tianbo","family":"Kang","sequence":"first","affiliation":[{"name":"National Key Laboratory of Intelligent Tracking and Forecasting for Infectious Diseases, Engineering Research Center of Trusted Behavior Intelligence, Ministry of Education, Tianjin Key Laboratory of Intelligent Robotics, Institute of Robotics and Automatic Information System, Nankai University, Tianjin 300350, China"},{"name":"Institute of Intelligence Technology and Robotic Systems, Shenzhen Research Institute of Nankai University, Shenzhen 518083, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Jianpeng","family":"Zhang","sequence":"additional","affiliation":[{"name":"National Key Laboratory of Intelligent Tracking and Forecasting for Infectious Diseases, Engineering Research Center of Trusted Behavior Intelligence, Ministry of Education, Tianjin Key Laboratory of Intelligent Robotics, Institute of Robotics and Automatic Information System, Nankai University, Tianjin 300350, China"},{"name":"Institute of Intelligence Technology and Robotic Systems, Shenzhen Research Institute of Nankai University, Shenzhen 518083, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Xin","family":"Zhao","sequence":"additional","affiliation":[{"name":"National Key Laboratory of Intelligent Tracking and Forecasting for Infectious Diseases, Engineering Research Center of Trusted Behavior Intelligence, Ministry of Education, Tianjin Key Laboratory of Intelligent Robotics, Institute of Robotics and Automatic Information System, Nankai University, Tianjin 300350, China"},{"name":"Institute of Intelligence Technology and Robotic Systems, Shenzhen Research Institute of Nankai University, Shenzhen 518083, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-9179-9668","authenticated-orcid":false,"given":"Mingzhu","family":"Sun","sequence":"additional","affiliation":[{"name":"National Key Laboratory of Intelligent Tracking and Forecasting for Infectious Diseases, Engineering Research Center of Trusted Behavior Intelligence, Ministry of Education, Tianjin Key Laboratory of Intelligent Robotics, Institute of Robotics and Automatic Information System, Nankai University, Tianjin 300350, China"},{"name":"Institute of Intelligence Technology and Robotic Systems, Shenzhen Research Institute of Nankai University, Shenzhen 518083, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yunwang","family":"Zhang","sequence":"additional","affiliation":[{"name":"Research Center of Laser Fusion, China Academy of Engineering Physics, Mianyang 621000, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"1968","published-online":{"date-parts":[[2026,5,28]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","unstructured":"Zhang, J., Dai, X., Wu, W., and Du, K. (2023). Micro-Vision Based High-Precision Space Assembly Approach for Trans-Scale Micro-Device: The CFTA Example. Sensors, 23.","DOI":"10.3390\/s23010450"},{"key":"ref_2","doi-asserted-by":"crossref","unstructured":"Redmon, J., Divvala, S., Girshick, R., and Farhadi, A. (2016, January 27\u201330). You only look once: Unified, real-time object detection. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Las Vegas, NV, USA.","DOI":"10.1109\/CVPR.2016.91"},{"key":"ref_3","doi-asserted-by":"crossref","first-page":"1137","DOI":"10.1109\/TPAMI.2016.2577031","article-title":"Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks","volume":"39","author":"Ren","year":"2017","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_4","doi-asserted-by":"crossref","unstructured":"Lin, T.Y., Goyal, P., Girshick, R., He, K., and Doll\u00e1r, P. (2017). Focal Loss for Dense Object Detection. Proceedings of the 2017 IEEE International Conference on Computer Vision (ICCV), IEEE.","DOI":"10.1109\/ICCV.2017.324"},{"key":"ref_5","doi-asserted-by":"crossref","unstructured":"Sun, X., Kong, H., Meng, Y., and Yang, X. (2017). Small Target Vehicle Detection Algorithm Based on Improved YOLOv5s. Proceedings of the 2024 8th CAA International Conference on Vehicular Control and Intelligence (CVCI), IEEE.","DOI":"10.1109\/CVCI63518.2024.10830200"},{"key":"ref_6","doi-asserted-by":"crossref","unstructured":"Wang, X., Chen, K., Yang, W., Yu, L., Xing, Y., and Yu, H. (2024). FE-DeTr: Keypoint detection and tracking in low-quality image frames with events. Proceedings of the 2024 IEEE International Conference on Robotics and Automation (ICRA), IEEE.","DOI":"10.1109\/ICRA57147.2024.10610579"},{"key":"ref_7","doi-asserted-by":"crossref","unstructured":"Bailey, D., Chang, Y., and Le Moan, S. (2020). Analysing Arbitrary Curves from the Line Hough Transform. J. Imaging, 6.","DOI":"10.3390\/jimaging6040026"},{"key":"ref_8","unstructured":"Leutenegger, M., and Weber, M. (2021). Least-squares Fitting of Gaussian Spots on Graphics Processing Units. arXiv."},{"key":"ref_9","doi-asserted-by":"crossref","first-page":"381","DOI":"10.1145\/358669.358692","article-title":"Random sample consensus: A paradigm for model fitting with applications to image analysis and automated cartography","volume":"24","author":"Fischler","year":"1981","journal-title":"Commun. ACM"},{"key":"ref_10","doi-asserted-by":"crossref","first-page":"1147","DOI":"10.1109\/TPAMI.2008.279","article-title":"Very Fast Best-Fit Circular and Elliptical Boundaries by Chord Data","volume":"31","author":"Barwick","year":"2009","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_11","doi-asserted-by":"crossref","first-page":"129","DOI":"10.1016\/j.measurement.2006.07.016","article-title":"Robust and accurate fitting of geometrical primitives to image data of microstructures","volume":"40","author":"Nehse","year":"2007","journal-title":"Measurement"},{"key":"ref_12","doi-asserted-by":"crossref","unstructured":"Yusong, Q., Liaomo, Z., Beibei, L., Dongdong, J., and Yuting, Z. (2023). High-precision Alignment System Based on Sub-pixel Segmentation. 2023 9th International Conference on Computer and Communications (ICCC), IEEE.","DOI":"10.1109\/ICCC59590.2023.10507711"},{"key":"ref_13","doi-asserted-by":"crossref","unstructured":"Zhang, X., Cui, Q., Bao, Q., Yang, W., and Liao, Q. (2024). Geometry-Guided Diffusion Model with Masked Transformer for Robust Multi-View 3D Human Pose Estimation. Proceedings of the 32nd ACM International Conference on Multimedia, Melbourn, VIC, Australia, 28 October\u20131 November, 2024, Association for Computing Machinery.","DOI":"10.1145\/3664647.3681265"},{"key":"ref_14","first-page":"1444","article-title":"Unsupervised Domain Adaptation through Shape Modeling for Medical Image Segmentation","volume":"Volume 172","author":"Konukoglu","year":"2022","journal-title":"Proceedings of the 5th International Conference on Medical Imaging with Deep Learning"},{"key":"ref_15","doi-asserted-by":"crossref","first-page":"e70129","DOI":"10.1049\/ipr2.70129","article-title":"A Systematic Review on Cell Nucleus Instance Segmentation","volume":"19","author":"Chen","year":"2025","journal-title":"IET Image Process."},{"key":"ref_16","unstructured":"Zhou, T., Xia, W., Zhang, F., Chang, B., Wang, W., Yuan, Y., Konukoglu, E., and Cremers, D. (2024). Image Segmentation in Foundation Model Era: A Survey. arXiv."},{"key":"ref_17","doi-asserted-by":"crossref","unstructured":"Ronneberger, O., Fischer, P., and Brox, T. (2015). U-net: Convolutional networks for biomedical image segmentation. Medical Image Computing and Computer-Assisted Intervention\u2014MICCAI 2015, Springer.","DOI":"10.1007\/978-3-319-24574-4_28"},{"key":"ref_18","doi-asserted-by":"crossref","unstructured":"He, K., Gkioxari, G., Doll\u00e1r, P., and Girshick, R. (2017). Mask r-cnn. Proceedings of the IEEE International Conference on Computer Vision, IEEE.","DOI":"10.1109\/ICCV.2017.322"},{"key":"ref_19","doi-asserted-by":"crossref","unstructured":"Chen, L.C., Zhu, Y., Papandreou, G., Schroff, F., and Adam, H. (2018). Encoder-decoder with atrous separable convolution for semantic image segmentation. Computer Vision\u2014ECCV 2018, Proceedings of the 15th European Conference on Computer Vision (ECCV), Munich, Germany, 8\u201314 September 2018, Springer.","DOI":"10.1007\/978-3-030-01234-2_49"},{"key":"ref_20","doi-asserted-by":"crossref","unstructured":"Kweon, H., and Yoon, K.J. (2025, January 11\u201315). WISH: Weakly Supervised Instance Segmentation using Heterogeneous Labels. Proceedings of the Computer Vision and Pattern Recognition Conference, Nashville, TN, USA.","DOI":"10.1109\/CVPR52734.2025.02363"},{"key":"ref_21","doi-asserted-by":"crossref","unstructured":"Kirillov, A., Mintun, E., Ravi, N., Mao, H., Rolland, C., Gustafson, L., Xiao, T., Whitehead, S., Berg, A.C., and Lo, W.Y. (2023). Segment anything. Proceedings of the IEEE\/CVF International Conference on Computer Vision, IEEE.","DOI":"10.1109\/ICCV51070.2023.00371"},{"key":"ref_22","unstructured":"Ravi, N., Gabeur, V., Hu, Y.T., Hu, R., Ryali, C., Ma, T., Khedr, H., R\u00e4dle, R., Rolland, C., and Gustafson, L. (2025). SAM 2: Segment Anything in Images and Videos. International Conference on Learning Representations 2025, ICLR."},{"key":"ref_23","unstructured":"Ren, T., Liu, S., Zeng, A., Lin, J., Li, K., Cao, H., Chen, J., Huang, X., Chen, Y., and Yan, F. (2024). Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks. arXiv."},{"key":"ref_24","doi-asserted-by":"crossref","unstructured":"Shaban, A., Bansal, S., Liu, Z., Essa, I., and Boots, B. (2017). One-Shot Learning for Semantic Segmentation. arXiv.","DOI":"10.5244\/C.31.167"},{"key":"ref_25","unstructured":"Catalano, N., and Matteucci, M. (2023). Few Shot Semantic Segmentation: A Review of Methodologies and Open Challenges. arXiv."},{"key":"ref_26","doi-asserted-by":"crossref","unstructured":"Wang, X., Wang, W., Cao, Y., Shen, C., and Huang, T. (2023). Images Speak in Images: A Generalist Painter for In-Context Visual Learning. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, IEEE.","DOI":"10.1109\/CVPR52729.2023.00660"},{"key":"ref_27","doi-asserted-by":"crossref","unstructured":"Wang, X., Zhang, X., Cao, Y., Wang, W., Shen, C., and Huang, T. (2023). SegGPT: Towards Segmenting Everything in Context. Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV), IEEE.","DOI":"10.1109\/ICCV51070.2023.00110"},{"key":"ref_28","doi-asserted-by":"crossref","first-page":"3365","DOI":"10.1109\/TPAMI.2020.2982166","article-title":"Deep Learning for Image Super-resolution: A Survey","volume":"43","author":"Wang","year":"2021","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_29","doi-asserted-by":"crossref","first-page":"2480","DOI":"10.1109\/TPAMI.2020.2968521","article-title":"Residual Dense Network for Image Restoration","volume":"43","author":"Zhang","year":"2021","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_30","doi-asserted-by":"crossref","first-page":"295","DOI":"10.1109\/TPAMI.2015.2439281","article-title":"Image Super-Resolution Using Deep Convolutional Networks","volume":"38","author":"Dong","year":"2016","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_31","first-page":"5602620","article-title":"MSMA-Net: An Infrared Small Target Detection Network by Multiscale Super-Resolution Enhancement and Multilevel Attention Fusion","volume":"62","author":"Ma","year":"2023","journal-title":"IEEE Trans. Geosci. Remote Sens."},{"key":"ref_32","unstructured":"Zhang, C., Cho, J., Puspitasari, F.D., Zheng, S., Li, C., Qiao, Y., Kang, T., Shan, X., Zhang, C., and Qin, C. (2023). A Survey on Segment Anything Model (SAM): Vision Foundation Model Meets Prompt Engineering. arXiv."},{"key":"ref_33","unstructured":"Zhang, C., Liu, L., Cui, Y., Huang, G., Lin, W., Yang, Y., and Hu, Y. (2023). A Comprehensive Survey on Segment Anything Model for Vision and Beyond. arXiv."},{"key":"ref_34","unstructured":"Dosovitskiy, A., Beyer, L., Kolesnikov, A., Weissenborn, D., Zhai, X., Unterthiner, T., Dehghani, M., Minderer, M., Heigold, G., and Gelly, S. (2021). An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale. arXiv."},{"key":"ref_35","doi-asserted-by":"crossref","unstructured":"Liang, J., Cao, J., Sun, G., Zhang, K., Van Gool, L., and Timofte, R. (2021). SwinIR: Image Restoration Using Swin Transformer. Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV) Workshops, IEEE.","DOI":"10.1109\/ICCVW54120.2021.00210"},{"key":"ref_36","doi-asserted-by":"crossref","first-page":"11382","DOI":"10.1109\/TPAMI.2025.3600126","article-title":"Rotation Equivariant Arbitrary-scale Image Super-Resolution","volume":"47","author":"Xie","year":"2025","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_37","unstructured":"Liang, J., Zhang, J., Gu, S., Gool, L.V., and Timofte, R. (2022). Revisiting RCAN: Improved Training for Image Super-Resolution. arXiv."}],"container-title":["Journal of Imaging"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/2313-433X\/12\/6\/238\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,6,4]],"date-time":"2026-06-04T04:18:35Z","timestamp":1780546715000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/2313-433X\/12\/6\/238"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,5,28]]},"references-count":37,"journal-issue":{"issue":"6","published-online":{"date-parts":[[2026,6]]}},"alternative-id":["jimaging12060238"],"URL":"https:\/\/doi.org\/10.3390\/jimaging12060238","relation":{},"ISSN":["2313-433X"],"issn-type":[{"value":"2313-433X","type":"electronic"}],"subject":[],"published":{"date-parts":[[2026,5,28]]}}}