{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,4]],"date-time":"2026-06-04T20:33:55Z","timestamp":1780605235478,"version":"3.54.1"},"reference-count":52,"publisher":"Institution of Engineering and Technology (IET)","issue":"1","license":[{"start":{"date-parts":[[2025,5,22]],"date-time":"2025-05-22T00:00:00Z","timestamp":1747872000000},"content-version":"vor","delay-in-days":141,"URL":"http:\/\/creativecommons.org\/licenses\/by-nc-nd\/4.0\/"},{"start":{"date-parts":[[2025,1,1]],"date-time":"2025-01-01T00:00:00Z","timestamp":1735689600000},"content-version":"tdm","delay-in-days":0,"URL":"http:\/\/doi.wiley.com\/10.1002\/tdm_license_1.1"}],"content-domain":{"domain":["ietresearch.onlinelibrary.wiley.com"],"crossmark-restriction":true},"short-container-title":["IET Image Processing"],"published-print":{"date-parts":[[2025,1]]},"abstract":"<jats:title>ABSTRACT<\/jats:title>\n                  <jats:p>Accurate classification of remote sensing scene images is crucial for diverse applications, from environmental monitoring to urban planning. While convolutional neural networks (CNNs) have dramatically improved classification accuracy, challenges remain due to the complex distribution of small objects, varied spatial configurations, and intra\u2010class multimodality in remote sensing images. In this work, we make three key contributions to address these challenges. (1) We propose the adaptive channel and coordinate relational attention network (ACDR\u2010CRAFF), a novel multi\u2010scale feature fusion framework designed to enhance feature representation across scales. (2) We introduce two innovative modules: the adaptive channel dimensionality reduction (ACDR) module, which dynamically adjusts channel representations to retain essential low\u2010dimensional features, and the coordinate relational attention multi\u2010scale feature fusion (CRAFF) module, which effectively captures and transfers spatial information between feature levels. (3) By integrating ACDR and CRAFF, our model achieves a progressive fusion of local to global features, ensuring robust feature expressiveness at multiple scales. Experimental results on four widely used benchmark datasets demonstrate that ACDR\u2010CRAFF consistently outperforms several state\u2010of\u2010the\u2010art methods, achieving significant improvements in classification accuracy and setting a new benchmark for complex remote sensing scene classification tasks. These results underscore the effectiveness of our approach in addressing the limitations of existing methods and advancing the state of the art in remote sensing image analysis.<\/jats:p>","DOI":"10.1049\/ipr2.70112","type":"journal-article","created":{"date-parts":[[2025,5,23]],"date-time":"2025-05-23T00:00:58Z","timestamp":1747958458000},"update-policy":"https:\/\/doi.org\/10.1002\/crossmark_policy","source":"Crossref","is-referenced-by-count":4,"title":["ACDR\u2010CRAFF Net: A Multi\u2010Scale Network Based on Adaptive Channel and Coordinate Relational Attention Network for Remote Sensing Scene Classification"],"prefix":"10.1049","volume":"19","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-7569-628X","authenticated-orcid":false,"given":"Wei","family":"Dai","sequence":"first","affiliation":[{"name":"School of Computer Science and Engineering Tianjin University of Technology Tianjin China"},{"name":"Key Laboratory of Computer Vision and System Ministry of Education Tianjin University of Technology Tianjin China"},{"name":"Tianjin Key Laboratory of Intelligence Computing and Novel Software Technology Tianjin University of Technology Tianjin China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Haixia","family":"Xu","sequence":"additional","affiliation":[{"name":"School of Computer Science and Engineering Tianjin University of Technology Tianjin China"},{"name":"Key Laboratory of Computer Vision and System Ministry of Education Tianjin University of Technology Tianjin China"},{"name":"Tianjin Key Laboratory of Intelligence Computing and Novel Software Technology Tianjin University of Technology Tianjin China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Furong","family":"Shi","sequence":"additional","affiliation":[{"name":"School of Computer Science and Engineering Tianjin University of Technology Tianjin China"},{"name":"Key Laboratory of Computer Vision and System Ministry of Education Tianjin University of Technology Tianjin China"},{"name":"Tianjin Key Laboratory of Intelligence Computing and Novel Software Technology Tianjin University of Technology Tianjin China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Liming","family":"Yuan","sequence":"additional","affiliation":[{"name":"School of Computer Science and Engineering Tianjin University of Technology Tianjin China"},{"name":"Key Laboratory of Computer Vision and System Ministry of Education Tianjin University of Technology Tianjin China"},{"name":"Tianjin Key Laboratory of Intelligence Computing and Novel Software Technology Tianjin University of Technology Tianjin China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-2746-6049","authenticated-orcid":false,"given":"Xinyu","family":"Wang","sequence":"additional","affiliation":[{"name":"School of Computer Science and Engineering Tianjin University of Technology Tianjin China"},{"name":"Key Laboratory of Computer Vision and System Ministry of Education Tianjin University of Technology Tianjin China"},{"name":"Tianjin Key Laboratory of Intelligence Computing and Novel Software Technology Tianjin University of Technology Tianjin China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0001-3702-0000","authenticated-orcid":false,"given":"Xianbin","family":"Wen","sequence":"additional","affiliation":[{"name":"School of Computer Science and Engineering Tianjin University of Technology Tianjin China"},{"name":"Key Laboratory of Computer Vision and System Ministry of Education Tianjin University of Technology Tianjin China"},{"name":"Tianjin Key Laboratory of Intelligence Computing and Novel Software Technology Tianjin University of Technology Tianjin China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"265","published-online":{"date-parts":[[2025,5,22]]},"reference":[{"key":"e_1_2_9_2_1","doi-asserted-by":"publisher","DOI":"10.1109\/JSTARS.2021.3117857"},{"key":"e_1_2_9_3_1","doi-asserted-by":"publisher","DOI":"10.1109\/JSTARS.2020.3005403"},{"key":"e_1_2_9_4_1","doi-asserted-by":"publisher","DOI":"10.1117\/1.JRS.16.044510"},{"key":"e_1_2_9_5_1","doi-asserted-by":"crossref","unstructured":"J.Hu G. S.Xia F.Hu H.Sun andL.Zhang \u201cA Comparative Study of Sampling Analysis in Scene Classification of High\u2010Resolution Remote Sensing Imagery \u201d in2015 IEEE International Geoscience and Remote Sensing Symposium (IGARSS)(IEEE 2015) 2389\u20132392.","DOI":"10.1109\/IGARSS.2015.7326290"},{"key":"e_1_2_9_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/83.817602"},{"key":"e_1_2_9_7_1","doi-asserted-by":"crossref","unstructured":"N.DalalandB.Triggs \u201cHistograms of Oriented Gradients for Human Detection \u201d in2005 IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR'05) Vol.1(IEEE 2005) 886\u2013893.","DOI":"10.1109\/CVPR.2005.177"},{"key":"e_1_2_9_8_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2002.1017623"},{"key":"e_1_2_9_9_1","doi-asserted-by":"crossref","unstructured":"Y.YangandS.Newsam \u201cBag\u2010of\u2010Visual\u2010Words and Spatial Extensions for Land\u2010Use Classification \u201d inProceedings of the 18th SIGSPATIAL International Conference on Advances in Geographic Information Systems(ACM 2010) 270\u2013279.","DOI":"10.1145\/1869790.1869829"},{"key":"e_1_2_9_10_1","doi-asserted-by":"crossref","unstructured":"J.Wang J.Yang K.Yu F.Lv T.Huang andY.Gong \u201cLocality\u2010Constrained Linear Coding for Image Classification \u201d in2010 IEEE Computer Society Conference on Computer Vision and Pattern Recognition(IEEE 2010) 3360\u20133367.","DOI":"10.1109\/CVPR.2010.5540018"},{"key":"e_1_2_9_11_1","doi-asserted-by":"crossref","unstructured":"S.Lazebnik C.Schmid andJ.Ponce \u201cBeyond Bags of Features: Spatial Pyramid Matching for Recognizing Natural Scene Categories \u201d in2006 IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR'06) Vol.2(IEEE 2006) 2169\u20132178.","DOI":"10.1109\/CVPR.2006.68"},{"key":"e_1_2_9_12_1","doi-asserted-by":"crossref","unstructured":"F.Perronnin J.S\u00e1nchez andT.Mensink \u201cImproving the Fisher Kernel for Large\u2010Scale Image Classification \u201d inComputer Vision\u2010ECCV 2010: 11th European Conference on Computer Vision Heraklion Crete Greece September 5\u201011 2010 Proceedings Part IV 11(Springer 2010) 143\u2013156.","DOI":"10.1007\/978-3-642-15561-1_11"},{"key":"e_1_2_9_13_1","doi-asserted-by":"publisher","DOI":"10.1049\/ipr2.12836"},{"key":"e_1_2_9_14_1","doi-asserted-by":"publisher","DOI":"10.1109\/JSTARS.2024.3456854"},{"key":"e_1_2_9_15_1","doi-asserted-by":"publisher","DOI":"10.1109\/LGRS.2020.2970810"},{"key":"e_1_2_9_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/TCYB.2020.3029787"},{"key":"e_1_2_9_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/LGRS.2021.3070016"},{"key":"e_1_2_9_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/TCYB.2021.3064571"},{"key":"e_1_2_9_19_1","doi-asserted-by":"publisher","DOI":"10.1109\/JSTARS.2021.3135566"},{"key":"e_1_2_9_20_1","doi-asserted-by":"publisher","DOI":"10.1109\/TNNLS.2023.3245643"},{"key":"e_1_2_9_21_1","doi-asserted-by":"publisher","DOI":"10.1109\/JSTARS.2022.3229729"},{"key":"e_1_2_9_22_1","first-page":"1","article-title":"EMTCAL: Efficient Multiscale Transformer and Cross\u2010Level Attention Learning for Remote Sensing Scene Classification","volume":"60","author":"Tang X.","year":"2022","journal-title":"IEEE Transactions on Geoscience and Remote Sensing"},{"key":"e_1_2_9_23_1","doi-asserted-by":"publisher","DOI":"10.3390\/rs17010042"},{"key":"e_1_2_9_24_1","unstructured":"A.Dosovitskiy \u201cAn Image is Worth 16x16 Words: Transformers for Image Recognition at Scale \u201darXiv preprint arXiv:2010.11929(2020)."},{"key":"e_1_2_9_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/TGRS.2020.3030990"},{"key":"e_1_2_9_26_1","doi-asserted-by":"crossref","unstructured":"L.Yuan Y.Chen T.Wang et\u00a0al. \u201cTokens\u2010to\u2010token vit: Training Vision Transformers From Scratch on Imagenet \u201d inProceedings of the IEEE\/CVF International Conference on Computer Vision(IEEE 2021) 558\u2013567.","DOI":"10.1109\/ICCV48922.2021.00060"},{"key":"e_1_2_9_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/JSTARS.2022.3155665"},{"key":"e_1_2_9_28_1","doi-asserted-by":"publisher","DOI":"10.3390\/rs15112865"},{"key":"e_1_2_9_29_1","doi-asserted-by":"crossref","unstructured":"B.Heo S.Yun D.Han S.Chun J.Choe andS. J.Oh \u201cRethinking Spatial Dimensions of Vision Transformers \u201d inProceedings of the IEEE\/CVF International Conference on Computer Vision(IEEE 2021) 11936\u201311945.","DOI":"10.1109\/ICCV48922.2021.01172"},{"key":"e_1_2_9_30_1","unstructured":"Y.Hu G.Wen M.Luo D.Dai J.Ma andZ.Yu \u201cCompetitive Inner\u2010Imaging Squeeze and Excitation for Residual Network \u201darXiv preprint arXiv:1807.08920(2018)."},{"key":"e_1_2_9_31_1","doi-asserted-by":"crossref","unstructured":"J.Hu L.Shen andG.Sun \u201cSqueeze\u2010and\u2010excitation Networks \u201d inProceedings of the IEEE Conference on Computer Vision and Pattern Recognition(IEEE 2018) 7132\u20137141.","DOI":"10.1109\/CVPR.2018.00745"},{"key":"e_1_2_9_32_1","doi-asserted-by":"crossref","unstructured":"Q.Wang B.Wu P.Zhu P.Li W.Zuo andQ.Hu \u201cECA\u2010Net: Efficient Channel Attention for Deep Convolutional Neural Networks \u201d inProceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition(IEEE 2020) 11534\u201311542.","DOI":"10.1109\/CVPR42600.2020.01155"},{"key":"e_1_2_9_33_1","unstructured":"J.Park S.Woo J. L.Yee andI. S.Kweon \u201cBam: Bottleneck Attention Module \u201darXiv preprint arXiv:1807.06514(2018)."},{"key":"e_1_2_9_34_1","doi-asserted-by":"crossref","unstructured":"S.Woo J.Park J. Y.Lee andI. S.Kweon \u201cCbam: Convolutional Block Attention Module \u201d inProceedings of the European Conference on Computer Vision (ECCV)(Springer 2018) 3\u201319.","DOI":"10.1007\/978-3-030-01234-2_1"},{"key":"e_1_2_9_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/TGRS.2018.2864987"},{"key":"e_1_2_9_36_1","doi-asserted-by":"publisher","DOI":"10.3390\/rs14092042"},{"key":"e_1_2_9_37_1","doi-asserted-by":"crossref","unstructured":"W.Wang Y.Shi andX.Wang \u201cRMFFNet: A Reverse Multi\u2010Scale Feature Fusion Network for Remote Sensing Scene Classification \u201d in2024 International Joint Conference on Neural Networks (IJCNN)(IEEE 2024) 1\u20138.","DOI":"10.1109\/IJCNN60899.2024.10650171"},{"key":"e_1_2_9_38_1","doi-asserted-by":"publisher","DOI":"10.1109\/TGRS.2023.3235819"},{"key":"e_1_2_9_39_1","doi-asserted-by":"publisher","DOI":"10.1038\/s41598-024-73252-8"},{"key":"e_1_2_9_40_1","doi-asserted-by":"crossref","unstructured":"K.He X.Zhang S.Ren andJ.Sun \u201cDeep Residual Learning for Image Recognition \u201d inProceedings of the IEEE Conference on Computer Vision and Pattern Recognition(IEEE 2016) 770\u2013778.","DOI":"10.1109\/CVPR.2016.90"},{"key":"e_1_2_9_41_1","unstructured":"M.Lin \u201cNetwork in Network \u201darXiv preprint arXiv:1312.4400(2013)."},{"key":"e_1_2_9_42_1","doi-asserted-by":"publisher","DOI":"10.1109\/TGRS.2017.2685945"},{"key":"e_1_2_9_43_1","doi-asserted-by":"publisher","DOI":"10.3390\/rs15194804"},{"key":"e_1_2_9_44_1","doi-asserted-by":"publisher","DOI":"10.1109\/JPROC.2017.2675998"},{"key":"e_1_2_9_45_1","doi-asserted-by":"publisher","DOI":"10.1109\/TGRS.2017.2685945"},{"key":"e_1_2_9_46_1","doi-asserted-by":"publisher","DOI":"10.3390\/rs11050494"},{"key":"e_1_2_9_47_1","doi-asserted-by":"publisher","DOI":"10.1109\/TNNLS.2020.3042276"},{"key":"e_1_2_9_48_1","doi-asserted-by":"publisher","DOI":"10.1109\/TNNLS.2019.2920374"},{"key":"e_1_2_9_49_1","doi-asserted-by":"publisher","DOI":"10.1109\/TGRS.2020.3044655"},{"key":"e_1_2_9_50_1","doi-asserted-by":"publisher","DOI":"10.1109\/TGRS.2019.2931801"},{"key":"e_1_2_9_51_1","doi-asserted-by":"publisher","DOI":"10.1109\/JSTARS.2020.3006241"},{"key":"e_1_2_9_52_1","doi-asserted-by":"crossref","unstructured":"W.Wang E.Xie X.Li et\u00a0al. \u201cPyramid Vision Transformer: A Versatile Backbone for Dense Prediction Without Convolutions \u201d inProceedings of the IEEE\/CVF International Conference on Computer Vision(IEEE 2021) 568\u2013578.","DOI":"10.1109\/ICCV48922.2021.00061"},{"key":"e_1_2_9_53_1","doi-asserted-by":"crossref","unstructured":"R. R.Selvaraju M.Cogswell A.Das R.Vedantam D.Parikh andD.Batra \u201cGrad\u2010CAM: Visual Explanations From Deep Networks via Gradient\u2010Based Localization \u201d in2017 IEEE International Conference on Computer Vision (ICCV)(IEEE 2017) 618\u2013626.","DOI":"10.1109\/ICCV.2017.74"}],"container-title":["IET Image Processing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/ietresearch.onlinelibrary.wiley.com\/doi\/pdf\/10.1049\/ipr2.70112","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/ietresearch.onlinelibrary.wiley.com\/doi\/full-xml\/10.1049\/ipr2.70112","content-type":"application\/xml","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/ietresearch.onlinelibrary.wiley.com\/doi\/pdf\/10.1049\/ipr2.70112","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,5,8]],"date-time":"2026-05-08T23:49:51Z","timestamp":1778284191000},"score":1,"resource":{"primary":{"URL":"https:\/\/ietresearch.onlinelibrary.wiley.com\/doi\/10.1049\/ipr2.70112"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,1]]},"references-count":52,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2025,1]]}},"alternative-id":["10.1049\/ipr2.70112"],"URL":"https:\/\/doi.org\/10.1049\/ipr2.70112","archive":["Portico"],"relation":{},"ISSN":["1751-9659","1751-9667"],"issn-type":[{"value":"1751-9659","type":"print"},{"value":"1751-9667","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,1]]},"assertion":[{"value":"2025-01-18","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2025-05-13","order":2,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2025-05-22","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}],"article-number":"e70112"}}