{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,9]],"date-time":"2026-05-09T07:46:52Z","timestamp":1778312812215,"version":"3.51.4"},"reference-count":65,"publisher":"MDPI AG","issue":"7","license":[{"start":{"date-parts":[[2025,7,7]],"date-time":"2025-07-07T00:00:00Z","timestamp":1751846400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"DOI":"10.13039\/501100004735","name":"Natural Science Foundation of Hunan Province","doi-asserted-by":"publisher","award":["2025JJ70028"],"award-info":[{"award-number":["2025JJ70028"]}],"id":[{"id":"10.13039\/501100004735","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100004735","name":"Natural Science Foundation of Hunan Province","doi-asserted-by":"publisher","award":["2025JJ81178"],"award-info":[{"award-number":["2025JJ81178"]}],"id":[{"id":"10.13039\/501100004735","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100004735","name":"Natural Science Foundation of Hunan Province","doi-asserted-by":"publisher","award":["2024JJ9550"],"award-info":[{"award-number":["2024JJ9550"]}],"id":[{"id":"10.13039\/501100004735","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100004735","name":"Natural Science Foundation of Hunan Province","doi-asserted-by":"publisher","award":["24A0401"],"award-info":[{"award-number":["24A0401"]}],"id":[{"id":"10.13039\/501100004735","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100009377","name":"Scientific Research Project of Education Department of Hunan Province","doi-asserted-by":"publisher","award":["2025JJ70028"],"award-info":[{"award-number":["2025JJ70028"]}],"id":[{"id":"10.13039\/100009377","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100009377","name":"Scientific Research Project of Education Department of Hunan Province","doi-asserted-by":"publisher","award":["2025JJ81178"],"award-info":[{"award-number":["2025JJ81178"]}],"id":[{"id":"10.13039\/100009377","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100009377","name":"Scientific Research Project of Education Department of Hunan Province","doi-asserted-by":"publisher","award":["2024JJ9550"],"award-info":[{"award-number":["2024JJ9550"]}],"id":[{"id":"10.13039\/100009377","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100009377","name":"Scientific Research Project of Education Department of Hunan Province","doi-asserted-by":"publisher","award":["24A0401"],"award-info":[{"award-number":["24A0401"]}],"id":[{"id":"10.13039\/100009377","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["BDCC"],"abstract":"<jats:p>Text-based person search (TPS), a critical technology for security and surveillance, aims to retrieve target individuals from image galleries using textual descriptions. The existing methods face two challenges: (1) ambiguous attribute\u2013noun association (AANA), where syntactic ambiguities lead to incorrect associations between attributes and the intended nouns; and (2) textual noise and relevance imbalance (TNRI), where irrelevant or non-discriminative tokens (e.g., \u2018wearing\u2019) reduce the saliency of critical visual attributes in the textual description. To address these aspects, we propose the dependency-aware entity\u2013attribute alignment network (DEAAN), a novel framework that explicitly tackles AANA through dependency-guided attention and TNRI via adaptive token filtering. The DEAAN introduces two modules: (1) dependency-assisted implicit reasoning (DAIR) to resolve AANA through syntactic parsing, and (2) relevance-adaptive token selection (RATS) to suppress TNRI by learning token saliency. Experiments on CUHK-PEDES, ICFG-PEDES, and RSTPReid demonstrate state-of-the-art performance, with the DEAAN achieving a Rank-1 accuracy of 76.71% and an mAP of 69.07% on CUHK-PEDES, surpassing RDE by 0.77% in Rank-1 and 1.51% in mAP. Ablation studies reveal that DAIR and RATS individually improve Rank-1 by 2.54% and 3.42%, while their combination elevates the performance by 6.35%, validating their synergy. This work bridges structured linguistic analysis with adaptive feature selection, demonstrating practical robustness in surveillance-oriented TPS scenarios.<\/jats:p>","DOI":"10.3390\/bdcc9070182","type":"journal-article","created":{"date-parts":[[2025,7,7]],"date-time":"2025-07-07T08:58:23Z","timestamp":1751878703000},"page":"182","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":1,"title":["Dependency-Aware Entity\u2013Attribute Relationship Learning for Text-Based Person Search"],"prefix":"10.3390","volume":"9","author":[{"ORCID":"https:\/\/orcid.org\/0009-0008-8601-9189","authenticated-orcid":false,"given":"Wei","family":"Xia","sequence":"first","affiliation":[{"name":"School of Computer, Hunan University of Technology, Zhuzhou 412000, China"},{"name":"School of Information Engineering, Hunan Applied Technology University, Changde 415000, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Wenguang","family":"Gan","sequence":"additional","affiliation":[{"name":"School of Computer, Hunan University of Technology, Zhuzhou 412000, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-9509-0755","authenticated-orcid":false,"given":"Xinpan","family":"Yuan","sequence":"additional","affiliation":[{"name":"School of Computer, Hunan University of Technology, Zhuzhou 412000, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1968","published-online":{"date-parts":[[2025,7,7]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","unstructured":"Jiang, D., and Ye, M. (2023, January 18\u201322). Cross-modal implicit relation reasoning and aligning for text-to-image person retrieval. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Vancouver, BC, Canada.","DOI":"10.1109\/CVPR52729.2023.00273"},{"key":"ref_2","doi-asserted-by":"crossref","unstructured":"Li, S., Xiao, T., Li, H., Zhou, B., Yue, D., and Wang, X. (2017, January 21\u201326). Person search with natural language description. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.551"},{"key":"ref_3","doi-asserted-by":"crossref","first-page":"9315","DOI":"10.1109\/TMM.2023.3251104","article-title":"Refined knowledge transfer for language-based person search","volume":"25","author":"Wu","year":"2023","journal-title":"IEEE Trans. Multimed."},{"key":"ref_4","doi-asserted-by":"crossref","unstructured":"Qin, Y., Chen, Y., Peng, D., Peng, X., Zhou, J.T., and Hu, P. (2024, January 16\u201322). Noisy-correspondence learning for text-to-image person re-identification. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Seattle, WA, USA.","DOI":"10.1109\/CVPR52733.2024.02568"},{"key":"ref_5","doi-asserted-by":"crossref","first-page":"104912","DOI":"10.1016\/j.imavis.2024.104912","article-title":"Eesso: Exploiting extreme and smooth signals via omni-frequency learning for text-based person retrieval","volume":"142","author":"Xue","year":"2024","journal-title":"Image Vis. Comput."},{"key":"ref_6","doi-asserted-by":"crossref","first-page":"4281","DOI":"10.1109\/TMM.2023.3321504","article-title":"Multi-granularity matching transformer for text-based person search","volume":"26","author":"Bao","year":"2023","journal-title":"IEEE Trans. Multimed."},{"key":"ref_7","doi-asserted-by":"crossref","unstructured":"Li, Y., Xu, H., and Xiao, J. (2020). Hybrid attention network for language-based person search. Sensors, 20.","DOI":"10.3390\/s20185279"},{"key":"ref_8","doi-asserted-by":"crossref","first-page":"2238","DOI":"10.1109\/LSP.2022.3217682","article-title":"Joint token and feature alignment framework for text-based person search","volume":"29","author":"Li","year":"2022","journal-title":"IEEE Signal Process. Lett."},{"key":"ref_9","unstructured":"Eom, C., and Ham, B. (2019, January 8\u201314). Learning disentangled representation for robust person re-identification. Proceedings of the 33rd International Conference on Neural Information Processing Systems, Vancouver, BC, Canada."},{"key":"ref_10","doi-asserted-by":"crossref","unstructured":"Wang, Z., Hu, R., Yu, Y., Liang, C., and Huang, W. (2015, January 26\u201330). Multi-level fusion for person re-identification with incomplete marks. Proceedings of the 23rd ACM International Conference on Multimedia, Brisbane, Australia.","DOI":"10.1145\/2733373.2806400"},{"key":"ref_11","doi-asserted-by":"crossref","unstructured":"Shu, X., Wen, W., Wu, H., Chen, K., Song, Y., Qiao, R., Ren, B., and Wang, X. (2022, January 23\u201327). See finer, see more: Implicit modality alignment for text-based person retrieval. Proceedings of the European Conference on Computer Vision, Tel Aviv, Israel.","DOI":"10.1007\/978-3-031-25072-9_42"},{"key":"ref_12","doi-asserted-by":"crossref","unstructured":"Zhang, Y., and Lu, H. (2018, January 8\u201314). Deep cross-modal projection learning for image-text matching. Proceedings of the European Conference on Computer Vision (ECCV), Munich, Germany.","DOI":"10.1007\/978-3-030-01246-5_42"},{"key":"ref_13","doi-asserted-by":"crossref","unstructured":"Wu, Y., Yan, Z., Han, X., Li, G., Zou, C., and Cui, S. (2021, January 11\u201317). LapsCore: Language-guided person search via color reasoning. Proceedings of the IEEE\/CVF International Conference on Computer Vision, Montreal, BC, Canada.","DOI":"10.1109\/ICCV48922.2021.00165"},{"key":"ref_14","doi-asserted-by":"crossref","unstructured":"Jing, Y., Si, C., Wang, J., Wang, W., Wang, L., and Tan, T. (2020, January 7\u201312). Pose-guided multi-granularity attention network for text-based person search. Proceedings of the AAAI Conference on Artificial Intelligence, New York, NY, USA.","DOI":"10.1609\/aaai.v34i07.6777"},{"key":"ref_15","doi-asserted-by":"crossref","first-page":"5542","DOI":"10.1109\/TIP.2020.2984883","article-title":"Improving description-based person re-identification by multi-granularity image-text alignments","volume":"29","author":"Niu","year":"2020","journal-title":"IEEE Trans. Image Process."},{"key":"ref_16","doi-asserted-by":"crossref","unstructured":"Wang, C., Luo, Z., Lin, Y., and Li, S. (2021, January 19\u201327). Text-based person search via multi-granularity embedding learning. Proceedings of the IJCAI, Montreal, BC, Canada.","DOI":"10.24963\/ijcai.2021\/148"},{"key":"ref_17","doi-asserted-by":"crossref","unstructured":"Shao, Z., Zhang, X., Fang, M., Lin, Z., Wang, J., and Ding, C. (2022, January 10\u201314). Learning granularity-unified representations for text-to-image person re-identification. Proceedings of the 30th ACM International Conference on Multimedia, Lisboa, Portugal.","DOI":"10.1145\/3503161.3548028"},{"key":"ref_18","doi-asserted-by":"crossref","unstructured":"Bai, Y., Cao, M., Gao, D., Cao, Z., Chen, C., Fan, Z., Nie, L., and Zhang, M. (2023, January 19\u201325). RaSa: Relation and sensitivity aware representation learning for text-based person search. Proceedings of the Thirty-Second International Joint Conference on Artificial Intelligence, Macao, China.","DOI":"10.24963\/ijcai.2023\/62"},{"key":"ref_19","doi-asserted-by":"crossref","first-page":"6609","DOI":"10.1109\/TMM.2024.3355644","article-title":"Cross-Modal Adaptive Dual Association for Text-to-Image Person Retrieval","volume":"26","author":"Lin","year":"2024","journal-title":"IEEE Trans. Multimed."},{"key":"ref_20","doi-asserted-by":"crossref","first-page":"5745","DOI":"10.1109\/TIFS.2025.3574970","article-title":"Granularity-Aware Hyperbolic Representation for Text-based Person Search","volume":"20","author":"Qi","year":"2025","journal-title":"IEEE Trans. Inf. Forensics Secur."},{"key":"ref_21","doi-asserted-by":"crossref","first-page":"6032","DOI":"10.1109\/TIP.2023.3327924","article-title":"Clip-driven fine-grained text-image person re-identification","volume":"32","author":"Yan","year":"2023","journal-title":"IEEE Trans. Image Process."},{"key":"ref_22","unstructured":"Li, J., Li, D., Xiong, C., and Hoi, S. (2022, January 17\u201323). Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation. Proceedings of the International Conference on Machine Learning, PMLR, Baltimore, MD, USA."},{"key":"ref_23","first-page":"9694","article-title":"Align before fuse: Vision and language representation learning with momentum distillation","volume":"34","author":"Li","year":"2021","journal-title":"Adv. Neural Inf. Process. Syst."},{"key":"ref_24","unstructured":"Ding, Z., Ding, C., Shao, Z., and Tao, D. (2021). Semantically self-aligned network for text-to-image part-aware person re-identification. arXiv."},{"key":"ref_25","doi-asserted-by":"crossref","unstructured":"Zhu, A., Wang, Z., Li, Y., Wan, X., Jin, J., Wang, T., Hu, F., and Hua, G. (2021, January 20\u201324). Dssl: Deep surroundings-person separation learning for text-based person retrieval. Proceedings of the 29th ACM International Conference on Multimedia, Virtual.","DOI":"10.1145\/3474085.3475369"},{"key":"ref_26","unstructured":"Li, Y. (2025). Entity Alignment in Multi-Lingual, Temporal, and Probabilistic Knowledge Graphs. [Ph.D. Thesis, Swinburne University of Technology]."},{"key":"ref_27","doi-asserted-by":"crossref","first-page":"14698","DOI":"10.48084\/etasr.7524","article-title":"Tweet Prediction for Social Media using Machine Learning","volume":"14","author":"Fattah","year":"2024","journal-title":"Eng. Technol. Appl. Sci. Res."},{"key":"ref_28","unstructured":"Bai, Y., Wang, J., Cao, M., Chen, C., Cao, Z., Nie, L., and Zhang, M. (November, January 29). Text-based person search without parallel image-text data. Proceedings of the 31st ACM International Conference on Multimedia, Ottawa, ON, Canada."},{"key":"ref_29","unstructured":"Cao, M., Bai, Y., Zeng, Z., Ye, M., and Zhang, M. (2024, January 26\u201327). An empirical study of clip for text-based person search. Proceedings of the AAAI Conference on Artificial Intelligence, Vancouver, BC, Canada."},{"key":"ref_30","unstructured":"Li, S., Xu, X., Yang, Y., Shen, F., Mo, Y., Li, Y., and Shen, H.T. (November, January 29). DCEL: Deep cross-modal evidential learning for text-based person retrieval. Proceedings of the 31st ACM International Conference on Multimedia, Ottawa, ON, Canada."},{"key":"ref_31","unstructured":"Ma, Y., Sun, X., Ji, J., Jiang, G., Zhuang, W., and Ji, R. (November, January 29). Beat: Bi-directional one-to-many embedding alignment for text-based person retrieval. Proceedings of the 31st ACM International Conference on Multimedia, Ottawa, ON, Canada."},{"key":"ref_32","unstructured":"Shen, F., Shu, X., Du, X., and Tang, J. (November, January 29). Pedestrian-specific bipartite-aware similarity learning for text-based person retrieval. Proceedings of the 31st ACM International Conference on Multimedia, Ottawa, ON, Canada."},{"key":"ref_33","doi-asserted-by":"crossref","first-page":"110253","DOI":"10.1016\/j.knosys.2023.110253","article-title":"Text-based person search via local-relational-global fine grained alignment","volume":"262","author":"Zhou","year":"2023","journal-title":"Knowl.-Based Syst."},{"key":"ref_34","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/3383184","article-title":"Dual-path convolutional image-text embeddings with instance loss","volume":"16","author":"Zheng","year":"2020","journal-title":"ACM Trans. Multimed. Comput. Commun. Appl. (TOMM)"},{"key":"ref_35","unstructured":"Gao, C., Cai, G., Jiang, X., Zheng, F., Zhang, J., Gong, Y., Peng, P., Guo, X., and Sun, X. (2021). Contextual non-local alignment over full-scale representation for text-based person search. arXiv."},{"key":"ref_36","unstructured":"Faghri, F., Fleet, D.J., Kiros, J.R., and Fidler, S. (2017). Vse++: Improving visual-semantic embeddings with hard negatives. arXiv."},{"key":"ref_37","unstructured":"Radford, A., Kim, J.W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., and Clark, J. (2021, January 18\u201324). Learning transferable visual models from natural language supervision. Proceedings of the International Conference on Machine Learning, PMLR, Virtual."},{"key":"ref_38","doi-asserted-by":"crossref","unstructured":"Han, X., He, S., Zhang, L., and Xiang, T. (2021). Text-based person search with limited data. arXiv.","DOI":"10.5244\/C.35.10"},{"key":"ref_39","doi-asserted-by":"crossref","unstructured":"Yuenyong, S., and Wongpatikaseree, K. (2022). Improving natural language person description search from videos with language model fine-tuning and approximate nearest neighbor. Big Data Cogn. Comput., 6.","DOI":"10.3390\/bdcc6040136"},{"key":"ref_40","doi-asserted-by":"crossref","first-page":"2129","DOI":"10.3390\/electronics14112129","article-title":"A Globally Collaborative Multi-View k-Means Clustering","volume":"14","author":"Sinaga","year":"2025","journal-title":"Electronics"},{"key":"ref_41","doi-asserted-by":"crossref","unstructured":"Suo, W., Sun, M., Niu, K., Gao, Y., Wang, P., Zhang, Y., and Wu, Q. (2022, January 23\u201327). A Simple and Robust Correlation Filtering Method for Text-Based Person Search. Proceedings of the European Conference on Computer Vision, Tel Aviv, Israel.","DOI":"10.1007\/978-3-031-19833-5_42"},{"key":"ref_42","doi-asserted-by":"crossref","first-page":"2881","DOI":"10.1109\/TCSVT.2024.3487908","article-title":"Cross-modal Uncertainty Modeling with Diffusion-based Refinement for Text-based Person Retrieval","volume":"35","author":"Li","year":"2024","journal-title":"IEEE Trans. Circuits Syst. Video Technol."},{"key":"ref_43","doi-asserted-by":"crossref","unstructured":"He, C., Li, S., Wang, Z., Shen, F., Yang, Y., and Xu, X. (2024, January 15\u201319). Diverse Embedding Modeling with Adaptive Noise Filter for Text-based Person Retrieval. Proceedings of the 2024 IEEE International Conference on Multimedia and Expo (ICME), Niagara Falls, ON, Canada.","DOI":"10.1109\/ICME57554.2024.10688112"},{"key":"ref_44","doi-asserted-by":"crossref","first-page":"110481","DOI":"10.1016\/j.patcog.2024.110481","article-title":"Text-based person search via cross-modal alignment learning","volume":"152","author":"Ke","year":"2024","journal-title":"Pattern Recognit."},{"key":"ref_45","first-page":"1","article-title":"Attention is all you need","volume":"30","author":"Vaswani","year":"2017","journal-title":"Adv. Neural Inf. Process. Syst."},{"key":"ref_46","doi-asserted-by":"crossref","unstructured":"Zhang, M., Li, Z., Fu, G., and Zhang, M. (2019). Syntax-enhanced neural machine translation with syntax-aware word representations. arXiv.","DOI":"10.18653\/v1\/N19-1118"},{"key":"ref_47","first-page":"3536","article-title":"Linguistic binding in diffusion models: Enhancing attribute correspondence through attention map alignment","volume":"36","author":"Rassin","year":"2023","journal-title":"Adv. Neural Inf. Process. Syst."},{"key":"ref_48","doi-asserted-by":"crossref","first-page":"2988","DOI":"10.1109\/TASLP.2023.3301214","article-title":"Syntax-aware data augmentation for neural machine translation","volume":"31","author":"Duan","year":"2023","journal-title":"IEEE\/ACM Trans. Audio Speech Lang. Process."},{"key":"ref_49","doi-asserted-by":"crossref","unstructured":"Zeng, P., Gao, L., Lyu, X., Jing, S., and Song, J. (2021, January 20\u201324). Conceptual and syntactical cross-modal alignment with cross-level consistency for image-text matching. Proceedings of the 29th ACM International Conference on Multimedia, Virtual.","DOI":"10.1145\/3474085.3475380"},{"key":"ref_50","doi-asserted-by":"crossref","unstructured":"Bugliarello, E., and Okazaki, N. (2020, January 5\u201310). Enhancing Machine Translation with Dependency-Aware Self-Attention. Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics, Online.","DOI":"10.18653\/v1\/2020.acl-main.147"},{"key":"ref_51","doi-asserted-by":"crossref","unstructured":"Li, Z., Zhou, Q., Li, C., Xu, K., and Cao, Y. (2021, January 1\u20136). Improving BERT with Syntax-aware Local Attention. Proceedings of the Findings of the Association for Computational Linguistics: ACL-IJCNLP 2021, Online.","DOI":"10.18653\/v1\/2021.findings-acl.57"},{"key":"ref_52","doi-asserted-by":"crossref","unstructured":"Xie, Y., Zhu, Z., Cheng, X., Huang, Z., and Chen, D. (2023, January 6\u201310). Syntax matters: Towards spoken language understanding via syntax-aware attention. Proceedings of the Findings of the Association for Computational Linguistics: EMNLP 2023, Singapore.","DOI":"10.18653\/v1\/2023.findings-emnlp.794"},{"key":"ref_53","doi-asserted-by":"crossref","unstructured":"Xu, Z., Guo, D., Tang, D., Su, Q., Shou, L., Gong, M., Zhong, W., Quan, X., Jiang, D., and Duan, N. (2021, January 1\u20136). Syntax-Enhanced Pre-trained Model. Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers), Online.","DOI":"10.18653\/v1\/2021.acl-long.420"},{"key":"ref_54","doi-asserted-by":"crossref","unstructured":"Honnibal, M., and Johnson, M. (2015, January 17\u201321). An improved non-monotonic transition system for dependency parsing. Proceedings of the Conference on Empirical Methods in Natural Language Processing, EMNLP 2015, Lisbon, Portugal.","DOI":"10.18653\/v1\/D15-1162"},{"key":"ref_55","doi-asserted-by":"crossref","unstructured":"Sennrich, R., Haddow, B., and Birch, A. (2016, January 7\u201312). Neural Machine Translation of Rare Words with Subword Units. Proceedings of the 54th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), Berlin, Germany.","DOI":"10.18653\/v1\/P16-1162"},{"key":"ref_56","doi-asserted-by":"crossref","unstructured":"Zhu, H., Ke, W., Li, D., Liu, J., Tian, L., and Shan, Y. (2022, January 21\u201324). Dual cross-attention learning for fine-grained visual categorization and object re-identification. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, New Orleans, LA, USA.","DOI":"10.1109\/CVPR52688.2022.00465"},{"key":"ref_57","doi-asserted-by":"crossref","unstructured":"He, K., Zhang, X., Ren, S., and Sun, J. (2016, January 27\u201330). Deep residual learning for image recognition. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Las Vegas, NV, USA.","DOI":"10.1109\/CVPR.2016.90"},{"key":"ref_58","unstructured":"Kingma, D.P., and Ba, J. (2014). Adam: A method for stochastic optimization. arXiv."},{"key":"ref_59","doi-asserted-by":"crossref","first-page":"171","DOI":"10.1016\/j.neucom.2022.04.081","article-title":"TIPCB: A simple but effective part-based convolutional baseline for text-based person search","volume":"494","author":"Chen","year":"2022","journal-title":"Neurocomputing"},{"key":"ref_60","doi-asserted-by":"crossref","unstructured":"Shao, Z., Zhang, X., Ding, C., Wang, J., and Wang, J. (2023, January 2\u20133). Unified pre-training with pseudo texts for text-to-image person re-identification. Proceedings of the IEEE\/CVF International Conference on Computer Vision, Paris, France.","DOI":"10.1109\/ICCV51070.2023.01026"},{"key":"ref_61","doi-asserted-by":"crossref","first-page":"163","DOI":"10.1109\/TIP.2023.3337653","article-title":"VGSG: Vision-guided semantic-group network for text-based person search","volume":"33","author":"He","year":"2023","journal-title":"IEEE Trans. Image Process."},{"key":"ref_62","doi-asserted-by":"crossref","unstructured":"Selvaraju, R.R., Cogswell, M., Das, A., Vedantam, R., Parikh, D., and Batra, D. (2017, January 22\u201329). Grad-cam: Visual explanations from deep networks via gradient-based localization. Proceedings of the IEEE International Conference on Computer Vision, Venice, Italy.","DOI":"10.1109\/ICCV.2017.74"},{"key":"ref_63","doi-asserted-by":"crossref","first-page":"14572","DOI":"10.48084\/etasr.7480","article-title":"Customer churn prediction for telecommunication companies using machine learning and ensemble methods","volume":"14","author":"Alotaibi","year":"2024","journal-title":"Eng. Technol. Appl. Sci. Res."},{"key":"ref_64","doi-asserted-by":"crossref","unstructured":"Wei, L., Zhang, S., Gao, W., and Tian, Q. (2018, January 18\u201323). Person transfer gan to bridge domain gap for person re-identification. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00016"},{"key":"ref_65","doi-asserted-by":"crossref","unstructured":"Wang, Z., Fang, Z., Wang, J., and Yang, Y. (2020, January 23\u201328). Vitaa: Visual-textual attributes alignment in person search by natural language. Proceedings of the Computer Vision\u2013ECCV 2020: 16th European Conference, Glasgow, UK. Proceedings, Part XII 16.","DOI":"10.1007\/978-3-030-58610-2_24"}],"container-title":["Big Data and Cognitive Computing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/2504-2289\/9\/7\/182\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,9]],"date-time":"2025-10-09T18:05:51Z","timestamp":1760033151000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/2504-2289\/9\/7\/182"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,7,7]]},"references-count":65,"journal-issue":{"issue":"7","published-online":{"date-parts":[[2025,7]]}},"alternative-id":["bdcc9070182"],"URL":"https:\/\/doi.org\/10.3390\/bdcc9070182","relation":{},"ISSN":["2504-2289"],"issn-type":[{"value":"2504-2289","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,7,7]]}}}