{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,28]],"date-time":"2026-06-28T06:12:01Z","timestamp":1782627121821,"version":"3.54.5"},"reference-count":41,"publisher":"MDPI AG","issue":"1","license":[{"start":{"date-parts":[[2025,1,2]],"date-time":"2025-01-02T00:00:00Z","timestamp":1735776000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"Research and Practice of Talent Cultivation Mode for Information Technology Innovation in Modern Industrial Colleges under the Background of New Engineering Education","award":["2024SJGLX0108"],"award-info":[{"award-number":["2024SJGLX0108"]}]},{"name":"Research and Practice of Talent Cultivation Mode for Information Technology Innovation in Modern Industrial Colleges under the Background of New Engineering Education","award":["82202270"],"award-info":[{"award-number":["82202270"]}]},{"name":"National Natural Science Foundation of China","award":["2024SJGLX0108"],"award-info":[{"award-number":["2024SJGLX0108"]}]},{"name":"National Natural Science Foundation of China","award":["82202270"],"award-info":[{"award-number":["82202270"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Symmetry"],"abstract":"<jats:p>Text-based person re-identification enables the retrieval of specific pedestrians from a large image library using textual descriptions, effectively addressing the issue of missing pedestrian images. The main challenges in this task are to learn discriminative image\u2013text features and achieve accurate cross-modal matching. Despite the potential of leveraging semantic information from pedestrian attributes, current methods have not yet fully harnessed this resource. To this end, we introduce a novel Text-based Dual-branch Person Re-identification Algorithm based on the Deep Attribute Information Mining (DAIM) network. Our approach employs a Masked Language Modeling (MLM) module to learn cross-modal attribute alignments through mask language modeling, and an Implicit Relational Prompt (IRP) module to extract relational cues between pedestrian attributes using tailored prompt templates. Furthermore, drawing inspiration from feature fusion techniques, we developed a Symmetry Semantic Feature Fusion (SSF) module that utilizes symmetric relationships between attributes to enhance the integration of information from different modes, aiming to capture comprehensive features and facilitate efficient cross-modal interactions. We evaluated our method using three benchmark datasets, CUHK-PEDES, ICFG-PEDES, and RSTPReid, and the results demonstrated Rank-1 accuracy rates of 78.17%, 69.47%, and 68.30%, respectively. These results indicate a significant enhancement in pedestrian retrieval accuracy, thereby validating the efficacy of our proposed approach.<\/jats:p>","DOI":"10.3390\/sym17010064","type":"journal-article","created":{"date-parts":[[2025,1,2]],"date-time":"2025-01-02T10:32:26Z","timestamp":1735813946000},"page":"64","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":1,"title":["A Text-Based Dual-Branch Person Re-Identification Algorithm Based on the Deep Attribute Information Mining Network"],"prefix":"10.3390","volume":"17","author":[{"given":"Ke","family":"Han","sequence":"first","affiliation":[{"name":"School of Information Engineering, North China University of Water Resources and Electric Power, Zhengzhou 450046, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Xiyan","family":"Zhang","sequence":"additional","affiliation":[{"name":"School of Information Engineering, North China University of Water Resources and Electric Power, Zhengzhou 450046, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Wenlong","family":"Xu","sequence":"additional","affiliation":[{"name":"School of Information Engineering, North China University of Water Resources and Electric Power, Zhengzhou 450046, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Long","family":"Jin","sequence":"additional","affiliation":[{"name":"School of Information Engineering, North China University of Water Resources and Electric Power, Zhengzhou 450046, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"1968","published-online":{"date-parts":[[2025,1,2]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","first-page":"2872","DOI":"10.1109\/TPAMI.2021.3054775","article-title":"Deep learning for person re-identification: A survey and outlook","volume":"44","author":"Ye","year":"2021","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_2","doi-asserted-by":"crossref","unstructured":"Aggarwal, S., Radhakrishnan, V.B., and Chakraborty, A. (2020, January 1\u20135). Text-based person search via attribute-aided matching. Proceedings of the IEEE\/CVF Winter Conference on Applications of Computer Vision, Snowmass Village, CO, USA.","DOI":"10.1109\/WACV45572.2020.9093640"},{"key":"ref_3","doi-asserted-by":"crossref","unstructured":"Cheng, Z., Fan, H., Wang, Q., Liu, S., and Tang, Y. (2023). Dual-Stage Attribute Embedding and Modality Consistency Learning-Based Visible\u2013Infrared Person Re-Identification. Electronics, 12.","DOI":"10.3390\/electronics12244892"},{"key":"ref_4","doi-asserted-by":"crossref","unstructured":"Cui, X., Liang, Y., and Zhang, W. (2023). Dual-Branch Person Re-Identification Algorithm Based on Multi-Feature Representation. Electronics, 12.","DOI":"10.3390\/electronics12081869"},{"key":"ref_5","first-page":"647","article-title":"A cross-modal pedestrian Re-ID algorithm based on dual attribute information","volume":"48","author":"Chen","year":"2022","journal-title":"J. Beijing Univ. Aeronaut. Astronaut."},{"key":"ref_6","unstructured":"Gong, T., Du, G., Wang, J., Ding, Y., and Zhang, L. (November, January 29). Prototype-guided Cross-modal Completion and Alignment for Incomplete Text-based Person Re-identification. Proceedings of the 31st ACM International Conference on Multimedia, Ottawa, ON, Canada."},{"key":"ref_7","unstructured":"Li, H., Yang, S., Zhang, Y., Tao, D., and Yu, Z. (2023). Progressive Feature Mining and External Knowledge-Assisted Text-Pedestrian Image Retrieval. arXiv."},{"key":"ref_8","doi-asserted-by":"crossref","first-page":"17973","DOI":"10.1109\/TNNLS.2023.3310118","article-title":"Image-specific information suppression and implicit local alignment for text-based person search","volume":"35","author":"Yan","year":"2023","journal-title":"IEEE Trans. Neural Netw. Learn. Syst."},{"key":"ref_9","doi-asserted-by":"crossref","unstructured":"Qin, Y., Chen, Y., Peng, D., Peng, X., Zhou, J.T., and Hu, P. (2024, January 16\u201322). Noisy-correspondence learning for text-to-image person re-identification. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Seattle, WA, USA.","DOI":"10.1109\/CVPR52733.2024.02568"},{"key":"ref_10","unstructured":"Devlin, J. (2018). Bert: Pre-training of deep bidirectional transformers for language understanding. arXiv."},{"key":"ref_11","unstructured":"Liang, W., and Liang, Y. (2024). DrBERT: Unveiling the potential of masked language modeling decoder in BERT pretraining. arXiv."},{"key":"ref_12","doi-asserted-by":"crossref","unstructured":"Czinczoll, T., H\u00f6nes, C., Schall, M., and De Melo, G. (2024, January 11\u201316). NextLevelBERT: Masked language modeling with higher-level representations for long documents. Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), Bangkok, Thailand.","DOI":"10.18653\/v1\/2024.acl-long.256"},{"key":"ref_13","doi-asserted-by":"crossref","unstructured":"Jiang, D., and Ye, M. (2023, January 17\u201324). Cross-modal implicit relation reasoning and aligning for text-to-image person retrieval. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Vancouver, BC, Canada.","DOI":"10.1109\/CVPR52729.2023.00273"},{"key":"ref_14","doi-asserted-by":"crossref","unstructured":"Fujii, T., and Tarashima, S. (2023, January 1\u20136). Bilma: Bidirectional local-matching for text-based person re-identification. Proceedings of the IEEE\/CVF International Conference on Computer Vision, Paris, France.","DOI":"10.1109\/ICCVW60793.2023.00295"},{"key":"ref_15","doi-asserted-by":"crossref","unstructured":"Zhu, H., Huang, J.H., Rudinac, S., and Kanoulas, E. (2024, January 10\u201314). Enhancing Interactive Image Retrieval with Query Rewriting Using Large Language Models and Vision Language Models. Proceedings of the 2024 International Conference on Multimedia Retrieval, Phuket, Thailand.","DOI":"10.1145\/3652583.3658032"},{"key":"ref_16","unstructured":"Huang, J.H., Alfadly, M., Ghanem, B., and Worring, M. (2023). Improving visual question answering models through robustness analysis and in-context learning with a chain of basic questions. arXiv."},{"key":"ref_17","unstructured":"Yang, S., and Zhang, Y. (2024). MLLMReID: Multimodal Large Language Model-based Person Re-identification. arXiv."},{"key":"ref_18","doi-asserted-by":"crossref","unstructured":"Zhang, D., Yu, Y., Li, C., Dong, J., Su, D., Chu, C., and Yu, D. (2024). Mm-llms: Recent advances in multimodal large language models. arXiv.","DOI":"10.18653\/v1\/2024.findings-acl.738"},{"key":"ref_19","first-page":"8748","article-title":"Learning transferable visual models from natural language supervision","volume":"139","author":"Radford","year":"2021","journal-title":"Proc. Int. Conf. Mach. Learn. PMLR"},{"key":"ref_20","doi-asserted-by":"crossref","unstructured":"Liu, F., Kim, M., Ren, Z., and Liu, X. (2024, January 16\u201322). Distilling CLIP with Dual Guidance for Learning Discriminative Human Body Shape Representation. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Seattle, WA, USA.","DOI":"10.1109\/CVPR52733.2024.00032"},{"key":"ref_21","doi-asserted-by":"crossref","unstructured":"Li, X., Zhang, Z., Tan, X., Chen, C., Qu, Y., Xie, Y., and Ma, L. (2024, January 16\u201322). Promptad: Learning prompts with only normal samples for few-shot anomaly detection. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Seattle, WA, USA.","DOI":"10.1109\/CVPR52733.2024.01594"},{"key":"ref_22","doi-asserted-by":"crossref","unstructured":"Phan, V.M.H., Xie, Y., Qi, Y., Liu, L., Liu, L., Zhang, B., Liao, Z., Wu, Q., To, M.S., and Verjans, J.W. (2024, January 16\u201322). Decomposing Disease Descriptions for Enhanced Pathology Detection: A Multi-Aspect Vision-Language Pre-training Framework. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Seattle, WA, USA.","DOI":"10.1109\/CVPR52733.2024.01092"},{"key":"ref_23","unstructured":"Wang, G., Yu, F., Li, J., Jia, Q., and Ding, S. (2023). Exploiting the textual potential from vision-language pre-training for text-based person search. arXiv."},{"key":"ref_24","doi-asserted-by":"crossref","unstructured":"Yang, Z., Wu, D., Wu, C., Lin, Z., Gu, J., and Wang, W. (2024, January 16\u201322). A Pedestrian is Worth One Prompt: Towards Language Guidance Person Re-Identification. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Seattle, WA, USA.","DOI":"10.1109\/CVPR52733.2024.01642"},{"key":"ref_25","doi-asserted-by":"crossref","unstructured":"Zhai, Y., Zeng, Y., Huang, Z., Qin, Z., Jin, X., and Cao, D. (2024, January 20\u201327). Multi-Prompts Learning with Cross-Modal Alignment for Attribute-Based Person Re-identification. Proceedings of the AAAI Conference on Artificial Intelligence, Vancouver, BC, Canada.","DOI":"10.1609\/aaai.v38i7.28524"},{"key":"ref_26","unstructured":"Li, W., Tan, L., Dai, P., and Zhang, Y. (2024). Prompt Decoupling for Text-to-Image Person Re-identification. arXiv."},{"key":"ref_27","unstructured":"Dosovitskiy, A. (2020). An image is worth 16 \u00d7 16 words: Transformers for image recognition at scale. arXiv."},{"key":"ref_28","first-page":"1","article-title":"Pre-train, prompt, and predict: A systematic survey of promptingmethods in natural language processing","volume":"55","author":"Liu","year":"2023","journal-title":"ACM Comput. Surv."},{"key":"ref_29","doi-asserted-by":"crossref","first-page":"4257","DOI":"10.1109\/TCSVT.2023.3243725","article-title":"Cross on cross attention: Deep fusion transformer for image captioning","volume":"33","author":"Zhang","year":"2023","journal-title":"IEEE Trans. Circuits Syst. Video Technol."},{"key":"ref_30","unstructured":"Yu, X., Dong, N., Zhu, L., Peng, H., and Tao, D. (2024). CLIP-Driven Semantic Discovery Network for Visible-Infrared Person Re-Identification. arXiv."},{"key":"ref_31","doi-asserted-by":"crossref","unstructured":"Li, S., Xiao, T., Li, H., Zhou, B., Yue, D., and Wang, X. (2017, January 21\u201326). Person search with natural language description. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.551"},{"key":"ref_32","unstructured":"Ding, Z., Ding, C., Shao, Z., and Tao, D. (2021). Semantically self-aligned network for text-to-image part-aware person re-identification. arXiv."},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"Zhu, A., Wang, Z., Li, Y., Wan, X., Jin, J., Wang, T., Hu, F., and Hua, G. (2021, January 20\u201324). Dssl: Deep surroundings-person separation learning for text-based person retrieval. Proceedings of the 29th ACM International Conference on Multimedia, Chengdu, China.","DOI":"10.1145\/3474085.3475369"},{"key":"ref_34","doi-asserted-by":"crossref","first-page":"1384","DOI":"10.11834\/jig.220620","article-title":"Transformer network for cross-modal text-to-image person re-identification","volume":"28","author":"Jiang","year":"2023","journal-title":"J. Image Graph."},{"key":"ref_35","unstructured":"Kingma, D.P. (2014). Adam: A method for stochastic optimization. arXiv."},{"key":"ref_36","doi-asserted-by":"crossref","unstructured":"Bai, Y., Cao, M., Gao, D., Cao, Z., Chen, C., Fan, Z., Nie, L., and Zhang, M. (2023). Rasa: Relation and sensitivity aware representation learning for text-based person search. arXiv.","DOI":"10.24963\/ijcai.2023\/62"},{"key":"ref_37","doi-asserted-by":"crossref","unstructured":"Shu, X., Wen, W., Wu, H., Chen, K., Song, Y., Qiao, R., Ren, B., and Wang, X. (2022, January 23\u201327). See finer, see more: Implicit modality alignment for text-based person retrieval. Proceedings of the European Conference on Computer Vision, Tel Aviv, Israel.","DOI":"10.1007\/978-3-031-25072-9_42"},{"key":"ref_38","unstructured":"Cao, M., Bai, Y., Zeng, Z., Ye, M., and Zhang, M. (2024, January 20\u201327). An Empirical Study of CLIP for Text-Based Person Search. Proceedings of the AAAI Conference on Artificial Intelligence, Vancouver, BC, Canada."},{"key":"ref_39","doi-asserted-by":"crossref","unstructured":"Zuo, J., Zhou, H., Nie, Y., Zhang, F., Guo, T., Sang, N., Wang, Y., and Gao, C. (2024, January 16\u201322). UFineBench: Towards Text-based Person Retrieval with Ultra-fine Granularity. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Seattle, WA, USA.","DOI":"10.1109\/CVPR52733.2024.02078"},{"key":"ref_40","unstructured":"Yang, S., Zhou, Y., Zheng, Z., Wang, Y., Zhu, L., and Wu, Y. (November, January 29). Towards unified text-based person retrieval: A large-scale multi-attribute and language search benchmark. Proceedings of the 31st ACM International Conference on Multimedia, Ottawa, ON, Canada."},{"key":"ref_41","unstructured":"Sun, J., Zheng, Z., and Ding, G. (2024). From Data Deluge to Data Curation: A Filtering-WoRA Paradigm for Efficient Text-based Person Search. arXiv."}],"container-title":["Symmetry"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/2073-8994\/17\/1\/64\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,7]],"date-time":"2025-10-07T15:23:44Z","timestamp":1759850624000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/2073-8994\/17\/1\/64"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,1,2]]},"references-count":41,"journal-issue":{"issue":"1","published-online":{"date-parts":[[2025,1]]}},"alternative-id":["sym17010064"],"URL":"https:\/\/doi.org\/10.3390\/sym17010064","relation":{},"ISSN":["2073-8994"],"issn-type":[{"value":"2073-8994","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,1,2]]}}}