{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,14]],"date-time":"2026-04-14T00:42:14Z","timestamp":1776127334377,"version":"3.50.1"},"reference-count":28,"publisher":"MDPI AG","issue":"12","license":[{"start":{"date-parts":[[2023,12,2]],"date-time":"2023-12-02T00:00:00Z","timestamp":1701475200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"Jie Bang Gua Shuai\u2019 Science and Technology Major Project of Liaoning Province in 2022","award":["2022JH1\/10400025"],"award-info":[{"award-number":["2022JH1\/10400025"]}]},{"name":"Jie Bang Gua Shuai\u2019 Science and Technology Major Project of Liaoning Province in 2022","award":["N2216010"],"award-info":[{"award-number":["N2216010"]}]},{"name":"Fundamental Research Funds for the Central Universities of China","award":["2022JH1\/10400025"],"award-info":[{"award-number":["2022JH1\/10400025"]}]},{"name":"Fundamental Research Funds for the Central Universities of China","award":["N2216010"],"award-info":[{"award-number":["N2216010"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Symmetry"],"abstract":"<jats:p>Multi-camera video surveillance has been widely applied in crowd statistics and analysis in smart city scenarios. Most existing studies rely on appearance or motion features for cross-camera trajectory tracking, due to the changing asymmetric perspectives of multiple cameras and occlusions in crowded scenes, resulting in low accuracy and poor tracking performance. This paper proposes a tracking method that fuses appearance and motion features. An implicit social model is used to obtain motion features containing spatio-temporal information and social relations for trajectory prediction. The TransReID model is used to obtain appearance features for re-identification. Fused features are derived by integrating appearance features, spatio-temporal information and social relations. Based on the fused features, multi-round clustering is adopted to associate cross-camera objects. Exclusively employing robust pedestrian reidentification and trajectory prediction models, coupled with the real-time detector YOLOX, without any reliance on supplementary information, an IDF1 score of 70.64% is attained on typical datasets derived from AiCity2023.<\/jats:p>","DOI":"10.3390\/sym15122145","type":"journal-article","created":{"date-parts":[[2023,12,2]],"date-time":"2023-12-02T13:45:41Z","timestamp":1701524741000},"page":"2145","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":2,"title":["Cross-Camera Tracking Model and Method Based on Multi-Feature Fusion"],"prefix":"10.3390","volume":"15","author":[{"given":"Peng","family":"Zhang","sequence":"first","affiliation":[{"name":"School of Computer Science and Engineering, Northeastern University, Shenyang 110169, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Siqi","family":"Wang","sequence":"additional","affiliation":[{"name":"School of Computer Science and Engineering, Northeastern University, Shenyang 110169, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Wei","family":"Zhang","sequence":"additional","affiliation":[{"name":"School of Computer Science and Engineering, Northeastern University, Shenyang 110169, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Weimin","family":"Lei","sequence":"additional","affiliation":[{"name":"School of Computer Science and Engineering, Northeastern University, Shenyang 110169, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xinlei","family":"Zhao","sequence":"additional","affiliation":[{"name":"Shenyang Er Yi San Electronic Technology Co., Ltd., Shenyang 110023, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Qingyang","family":"Jing","sequence":"additional","affiliation":[{"name":"School of Computer Science and Engineering, Northeastern University, Shenyang 110169, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Mingxin","family":"Liu","sequence":"additional","affiliation":[{"name":"School of Computer Science and Engineering, Northeastern University, Shenyang 110169, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1968","published-online":{"date-parts":[[2023,12,2]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","first-page":"2540","DOI":"10.1609\/aaai.v36i3.20155","article-title":"Pose-Guided Feature Disentangling for Occluded Person Re-Identification Based on Transformer","volume":"36","author":"Wang","year":"2022","journal-title":"AAAI"},{"key":"ref_2","doi-asserted-by":"crossref","unstructured":"Somers, V., Vleeschouwer, C.D., and Alahi, A. (2023, January 2\u20137). Body Part-Based Representation Learning for Occluded Person Re-Identification. Proceedings of the 2023 IEEE\/CVF Winter Conference on Applications of Computer Vision (WACV), Waikoloa, HI, USA.","DOI":"10.1109\/WACV56688.2023.00166"},{"key":"ref_3","doi-asserted-by":"crossref","first-page":"17633","DOI":"10.1007\/s00521-022-07400-4","article-title":"Short Range Correlation Transformer for Occluded Person Re-Identification","volume":"34","author":"Zhao","year":"2022","journal-title":"Neural Comput. Appl."},{"key":"ref_4","first-page":"4894","article-title":"Feature Completion for Occluded Person Re-Identification","volume":"44","author":"Hou","year":"2021","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_5","unstructured":"Mohamed, A., Zhu, D., Vu, W., Elhoseiny, M., and Claudel, C. (2022). European Conference on Computer Vision, Springer."},{"key":"ref_6","doi-asserted-by":"crossref","unstructured":"Alahi, A., Goel, K., Ramanathan, V., Robicquet, A., Fei-Fei, L., and Savarese, S. (July, January 26). Social LSTM: Human Trajectory Prediction in Crowded Spaces. Proceedings of the 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Las Vegas, NV, USA.","DOI":"10.1109\/CVPR.2016.110"},{"key":"ref_7","doi-asserted-by":"crossref","unstructured":"Mohamed, A., Qian, K., Elhoseiny, M., and Claudel, C. (2020, January 13\u201319). Social-STGCNN: A Social Spatio-Temporal Graph Convolutional Neural Network for Human Trajectory Prediction. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Seattle, WA, USA.","DOI":"10.1109\/CVPR42600.2020.01443"},{"key":"ref_8","doi-asserted-by":"crossref","unstructured":"Rhinehart, N., Mcallister, R., Kitani, K., and Levine, S. (November, January 27). PRECOG: PREdiction Conditioned on Goals in Visual Multi-Agent Settings. Proceedings of the 2019 IEEE\/CVF International Conference on Computer Vision (ICCV), Seoul, Republic of Korea.","DOI":"10.1109\/ICCV.2019.00291"},{"key":"ref_9","doi-asserted-by":"crossref","unstructured":"Yuan, Y., Weng, X., Ou, Y., and Kitani, K. (2021, January 11\u201317). AgentFormer: Agent-Aware Transformers for Socio-Temporal Multi-Agent Forecasting. Proceedings of the 2021 IEEE\/CVF International Conference on Computer Vision (ICCV), Montreal, QC, Canada.","DOI":"10.1109\/ICCV48922.2021.00967"},{"key":"ref_10","doi-asserted-by":"crossref","unstructured":"Gupta, A., Johnson, J., Fei-Fei, L., Savarese, S., and Alahi, A. (2018, January 18\u201323). Social GAN: Socially Acceptable Trajectories with Generative Adversarial Networks. Proceedings of the 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00240"},{"key":"ref_11","doi-asserted-by":"crossref","unstructured":"Sadeghian, A., Kosaraju, V., Sadeghian, A., Hirose, N., Rezatofighi, H., and Savarese, S. (2019, January 15\u201320). SoPhie: An Attentive GAN for Predicting Paths Compliant to Social and Physical Constraints. Proceedings of the 2019 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00144"},{"key":"ref_12","unstructured":"Kosaraju, V., Sadeghian, A., Mart\u00edn-Mart\u00edn, R., Reid, I., Rezatofighi, H., and Savarese, S. (2019). Advances in Neural Information Processing Systems, Curran Associates, Inc."},{"key":"ref_13","unstructured":"He, S., Luo, H., Wang, P., Wang, F., Li, H., and Jiang, W. (2023, January 11\u201317). TransReID: Transformer-Based Object Re-Identification. Proceedings of the IEEE\/CVF International Conference on Computer Vision, Montreal, BC, Canada."},{"key":"ref_14","doi-asserted-by":"crossref","first-page":"14","DOI":"10.1109\/TDSC.2012.74","article-title":"SORT: A Self-ORganizing Trust Model for Peer-to-Peer Systems","volume":"10","author":"Can","year":"2013","journal-title":"IEEE Trans. Depend. Secur. Comput."},{"key":"ref_15","doi-asserted-by":"crossref","unstructured":"Wojke, N., Bewley, A., and Paulus, D. (2017, January 17\u201320). Simple Online and Realtime Tracking with a Deep Association Metric. Proceedings of the 2017 IEEE International Conference on Image Processing (ICIP), Beijing, China.","DOI":"10.1109\/ICIP.2017.8296962"},{"key":"ref_16","doi-asserted-by":"crossref","first-page":"3069","DOI":"10.1007\/s11263-021-01513-4","article-title":"FairMOT: On the Fairness of Detection and Re-Identification in Multiple Object Tracking","volume":"129","author":"Zhang","year":"2021","journal-title":"Int. J. Comput. Vis."},{"key":"ref_17","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1007\/978-3-031-20047-2_1","article-title":"ByteTrack: Multi-Object Tracking by Associating Every Detection Box","volume":"Volume 13682","author":"Avidan","year":"2022","journal-title":"Computer Vision\u2013ECCV 2022"},{"key":"ref_18","unstructured":"Aharon, N., Orfaig, R., and Bobrovsky, B.-Z. (2022). BoT-SORT: Robust Associations Multi-Pedestrian Tracking. arXiv."},{"key":"ref_19","unstructured":"Milos, S.S., Nemanja, I., and Srdan, S. (2021). Decentralized Consensus-Based Estimation and Target Tracking, Akademska misao."},{"key":"ref_20","unstructured":"You, Q., and Jiang, H. (2020). Real-Time 3D Deep Multi-Camera Tracking. arXiv."},{"key":"ref_21","doi-asserted-by":"crossref","unstructured":"Quach, K.G., Nguyen, P., Le, H., Truong, T.-D., Duong, C.N., Tran, M.-T., and Luu, K. (2021, January 20\u201325). DyGLIP: A Dynamic Graph Model with Link Prediction for Accurate Multi-Camera Multiple Object Tracking. Proceedings of the 2021 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Nashville, TN, USA.","DOI":"10.1109\/CVPR46437.2021.01357"},{"key":"ref_22","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1007\/978-3-030-58571-6_1","article-title":"Multiview Detection with Feature Perspective Transformation","volume":"Volume 12352","author":"Vedaldi","year":"2020","journal-title":"Computer Vision\u2013ECCV 2020"},{"key":"ref_23","doi-asserted-by":"crossref","unstructured":"Nguyen, D.M.H., Henschel, R., Rosenhahn, B., Sonntag, D., and Swoboda, P. (2022, January 18\u201324). LMGP: Lifted Multicut Meets Geometry Projections for Multi-Camera Multi-Object Tracking. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, New Orleans, LA, USA.","DOI":"10.1109\/CVPR52688.2022.00866"},{"key":"ref_24","unstructured":"Li, K., and Malik, J. (2018). Implicit Maximum Likelihood Estimation. arXiv."},{"key":"ref_25","unstructured":"Zheng, L., Shen, L., Tian, L., Wang, S., Wang, J., and Tian, Q. (2023, January 7\u201313). Scalable Person Re-Identification: A Benchmark. Proceedings of the IEEE International Conference on Computer Vision, Santiago, Chile."},{"key":"ref_26","doi-asserted-by":"crossref","unstructured":"Wei, L., Zhang, S., Gao, W., and Tian, Q. (2018, January 18\u201323). Person Transfer GAN to Bridge Domain Gap for Person Re-Identification. Proceedings of the 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00016"},{"key":"ref_27","unstructured":"Xiao, Q., Luo, H., and Zhang, C. (2017). Margin Sample Mining Loss: A Deep Learning Based Method for Person Re-Identification."},{"key":"ref_28","doi-asserted-by":"crossref","unstructured":"Jeon, Y., Tran, D.Q., Park, M., and Park, S. (2023, January 18\u201319). Leveraging Future Trajectory Prediction for Multi-Camera People Tracking. Proceedings of the 2023 IEEE\/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW), Vancouver, BC, Canada.","DOI":"10.1109\/CVPRW59228.2023.00570"}],"container-title":["Symmetry"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/2073-8994\/15\/12\/2145\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,10]],"date-time":"2025-10-10T21:36:40Z","timestamp":1760132200000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/2073-8994\/15\/12\/2145"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,12,2]]},"references-count":28,"journal-issue":{"issue":"12","published-online":{"date-parts":[[2023,12]]}},"alternative-id":["sym15122145"],"URL":"https:\/\/doi.org\/10.3390\/sym15122145","relation":{},"ISSN":["2073-8994"],"issn-type":[{"value":"2073-8994","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,12,2]]}}}