{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,7,30]],"date-time":"2025-07-30T14:09:39Z","timestamp":1753884579639,"version":"3.41.2"},"reference-count":31,"publisher":"World Scientific Pub Co Pte Ltd","issue":"07","funder":[{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["11801199"],"award-info":[{"award-number":["11801199"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]},{"name":"Natural Science Foundation of Anhui Provinc","award":["1908085QA30"],"award-info":[{"award-number":["1908085QA30"]}]},{"name":"Wannan Medical College key project research fund","award":["WK2023ZZD04"],"award-info":[{"award-number":["WK2023ZZD04"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["J CIRCUIT SYST COMP"],"published-print":{"date-parts":[[2025,5,15]]},"abstract":"<jats:p> Oral cancer is one of the most prevalent cancers globally, characterized by high rates of recurrence and metastasis, making early diagnosis crucial for improving patient survival. Accurate preoperative tumor-node-metastasis (TNM) staging is essential for determining effective treatment plans and surgical strategies. Although pathological examination remains the gold standard for TNM staging of oral cancer, imaging data can complement pathology, helping clinicians make more precise preoperative assessments and addressing the limitations of single-modality approaches. In this study, we propose a novel multimodal model for oral cancer staging that integrates preoperative CT images and whole-slide imaging (WSI). To overcome the heterogeneity between CT and WSI features, especially the lack of interaction between macrolevel CT features and microlevel pathological features, we introduce a CT-guided collaborative attention module. Specifically, CT features serve as queries, while patches of WSI are treated as key-value pairs. The corresponding keys are computed through a fully connected layer with trainable weights. The model architecture includes two separate pathways for feature extraction: CT features are extracted using a U-Net-based network, while WSI features are extracted with a multi-instance network utilizing a hierarchical attention mechanism. The CT-guided collaborative attention module facilitates interaction and fusion between these two modalities, resulting in a unified feature representation. This fused feature is then passed through a fully connected layer to produce the final TNM staging prediction for oral cancer. By combining both macro- and microlevel features, our model addresses the limitations of traditional single-modality methods, enabling more accurate tumor boundary delineation. Compared to existing techniques, our multimodal approach improves diagnostic accuracy and provides more reliable staging results. This advancement has the potential to significantly enhance early diagnosis and treatment strategies for oral cancer, offering a more comprehensive and precise method for staging the disease. <\/jats:p>","DOI":"10.1142\/s0218126625501737","type":"journal-article","created":{"date-parts":[[2024,12,27]],"date-time":"2024-12-27T07:17:06Z","timestamp":1735283826000},"source":"Crossref","is-referenced-by-count":0,"title":["A Multimodal Deep Learning Approach for Pre-Operative Tumor-Node-Metastasis Staging in Oral Cancer Using Computed Tomography Imaging and Pathological Slide Data"],"prefix":"10.1142","volume":"34","author":[{"given":"Huimin","family":"Jiang","sequence":"first","affiliation":[{"name":"Department of Medical Imaging, Wannan Medical College Wuhu, Anhui 241002, P.\u00a0R.\u00a0China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0005-1807-3368","authenticated-orcid":false,"given":"Liming","family":"Fang","sequence":"additional","affiliation":[{"name":"Department of Medical Imaging, Wannan Medical College Wuhu, Anhui 241002, P.\u00a0R.\u00a0China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yong","family":"Cui","sequence":"additional","affiliation":[{"name":"Oral and Maxillofacial Surgery, The First Affiliated Hospital of Wannan Medical College Wuhu, Anhui 241002, P.\u00a0R.\u00a0China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Anqi","family":"Xia","sequence":"additional","affiliation":[{"name":"Department of Medical Imaging, Wannan Medical College Wuhu, Anhui 241002, P.\u00a0R.\u00a0China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jing","family":"Wu","sequence":"additional","affiliation":[{"name":"Department of Imaging, The First Affiliated Hospital of Wannan Medical College Wuhu, Anhui 241002, P.\u00a0R.\u00a0China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jimin","family":"Chu","sequence":"additional","affiliation":[{"name":"Department of Medical Imaging, Wannan Medical College Wuhu, Anhui 241002, P.\u00a0R.\u00a0China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"219","published-online":{"date-parts":[[2025,2,26]]},"reference":[{"key":"S0218126625501737BIB001","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-32316-5"},{"key":"S0218126625501737BIB002","doi-asserted-by":"publisher","DOI":"10.1371\/journal.pone.0265950"},{"key":"S0218126625501737BIB003","doi-asserted-by":"publisher","DOI":"10.1016\/j.ijrobp.2022.03.033"},{"key":"S0218126625501737BIB004","volume-title":"AJCC Cancer Staging Manual","author":"Edge S. B.","year":"2018","edition":"8"},{"key":"S0218126625501737BIB005","volume-title":"TNM Classification of Malignant Tumors","author":"Brierley J.","year":"2017","edition":"8"},{"key":"S0218126625501737BIB006","doi-asserted-by":"publisher","DOI":"10.1002\/lary.27205"},{"key":"S0218126625501737BIB007","doi-asserted-by":"publisher","DOI":"10.1093\/ajcp\/aqy143"},{"key":"S0218126625501737BIB008","doi-asserted-by":"publisher","DOI":"10.1002\/lary.29537"},{"key":"S0218126625501737BIB009","doi-asserted-by":"publisher","DOI":"10.1002\/cncr.35020"},{"key":"S0218126625501737BIB010","doi-asserted-by":"publisher","DOI":"10.1038\/s43018-023-00694-w"},{"key":"S0218126625501737BIB011","doi-asserted-by":"publisher","DOI":"10.1016\/j.cgh.2024.02.007"},{"key":"S0218126625501737BIB012","first-page":"342","volume":"24","author":"Warin K.","year":"2022","journal-title":"PLoS One"},{"key":"S0218126625501737BIB013","doi-asserted-by":"publisher","DOI":"10.1016\/j.media.2017.08.006"},{"key":"S0218126625501737BIB014","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-24574-4_28"},{"key":"S0218126625501737BIB016","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP40776.2020.9053405"},{"first-page":"74","volume-title":"23rd Int. Conf. Medical Image Computing and Computer Assisted Intervention\u2013MICCAI 2020","author":"Xiang T.","key":"S0218126625501737BIB017"},{"key":"S0218126625501737BIB018","doi-asserted-by":"publisher","DOI":"10.1186\/s40708-023-00217-4"},{"key":"S0218126625501737BIB019","doi-asserted-by":"publisher","DOI":"10.1038\/s41592-020-01008-z"},{"key":"S0218126625501737BIB020","doi-asserted-by":"publisher","DOI":"10.1016\/j.compmedimag.2023.102303"},{"key":"S0218126625501737BIB022","doi-asserted-by":"publisher","DOI":"10.1109\/WACV51458.2022.00181"},{"key":"S0218126625501737BIB023","series-title":"Lecture Notes in Computer Science","doi-asserted-by":"crossref","first-page":"272","DOI":"10.1007\/978-3-031-08999-2_22","volume-title":"BrainLes 2021","volume":"12962","author":"Hatamizadeh A.","year":"2021"},{"key":"S0218126625501737BIB024","doi-asserted-by":"publisher","DOI":"10.1016\/j.artmed.2022.102475"},{"first-page":"1","volume-title":"IEEE Int. Conf. Acoustics, Speech and Signal Processing (ICASSP)","author":"Chatzianastasis M.","key":"S0218126625501737BIB025"},{"key":"S0218126625501737BIB026","doi-asserted-by":"publisher","DOI":"10.1109\/MWC.001.2000374"},{"key":"S0218126625501737BIB027","doi-asserted-by":"publisher","DOI":"10.1007\/s00521-021-06219-9"},{"key":"S0218126625501737BIB028","doi-asserted-by":"publisher","DOI":"10.1109\/JBHI.2024.3422180"},{"first-page":"5802","volume-title":"Proc. IEEE\/CVF Conf. Computer Vision and Pattern Recognition (CVPR)","author":"Liu J.","key":"S0218126625501737BIB029"},{"key":"S0218126625501737BIB030","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2018.2887342"},{"key":"S0218126625501737BIB031","first-page":"8026","volume":"32","author":"Paszke A.","year":"2019","journal-title":"Adv. Neural Inf. Process. Syst."},{"key":"S0218126625501737BIB033","doi-asserted-by":"publisher","DOI":"10.1109\/TNSE.2022.3190765"},{"key":"S0218126625501737BIB034","doi-asserted-by":"publisher","DOI":"10.1109\/TCSS.2022.3222682"}],"container-title":["Journal of Circuits, Systems and Computers"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.worldscientific.com\/doi\/pdf\/10.1142\/S0218126625501737","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,4,26]],"date-time":"2025-04-26T05:47:17Z","timestamp":1745646437000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.worldscientific.com\/doi\/10.1142\/S0218126625501737"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,2,26]]},"references-count":31,"journal-issue":{"issue":"07","published-print":{"date-parts":[[2025,5,15]]}},"alternative-id":["10.1142\/S0218126625501737"],"URL":"https:\/\/doi.org\/10.1142\/s0218126625501737","relation":{},"ISSN":["0218-1266","1793-6454"],"issn-type":[{"type":"print","value":"0218-1266"},{"type":"electronic","value":"1793-6454"}],"subject":[],"published":{"date-parts":[[2025,2,26]]},"article-number":"2550173"}}