{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,7,30]],"date-time":"2025-07-30T14:17:31Z","timestamp":1753885051032,"version":"3.41.2"},"reference-count":9,"publisher":"World Scientific Pub Co Pte Ltd","issue":"01","funder":[{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"crossref","award":["61962019","61961017"],"award-info":[{"award-number":["61962019","61961017"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"crossref"}]},{"name":"Enshi Prefecture","award":["D20220005"],"award-info":[{"award-number":["D20220005"]}]},{"name":"National Cultural and Tourism Science and Technology innovation project","award":["2021064"],"award-info":[{"award-number":["2021064"]}]},{"name":"Hubei Engineering Research Center of Selenium Food Nutrition and Health Intelligent Technology, the Hubei Minzu University","award":["PT082004","PT082105"],"award-info":[{"award-number":["PT082004","PT082105"]}]},{"name":"Hubei Soft Science Research Program of China","award":["2022EDA065"],"award-info":[{"award-number":["2022EDA065"]}]},{"name":"Hubei Minzu University of College of Intelligent Systems Science and Engineering","award":["ZYK2022010"],"award-info":[{"award-number":["ZYK2022010"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Int. J. Image Grap."],"published-print":{"date-parts":[[2025,1]]},"abstract":"<jats:p> Vision transformers are deep neural networks applied to image classification based on a self-attention mechanism and can process data in parallel. Aiming at the structural loss of Vision transformers, this paper combines ConViT and Convolutional Neural Network (CNN) and proposes a new model Convolution Meet Vision Transformers (CMVT). This model adds a convolution module to the ConViT network to solve the structural loss of the transformer. By adding hierarchical data representation, the ability to gradually extract more image classification features is improved. We have conducted comparative experiments on multiple dataset, and all of them have been enhanced to improve the efficiency and performance of the model. <\/jats:p>","DOI":"10.1142\/s0219467824500608","type":"journal-article","created":{"date-parts":[[2024,5,6]],"date-time":"2024-05-06T09:51:31Z","timestamp":1714989091000},"source":"Crossref","is-referenced-by-count":0,"title":["CMVT: ConVit Transformer Network Recombined with Convolutional Layer"],"prefix":"10.1142","volume":"25","author":[{"given":"Chunxia","family":"Mao","sequence":"first","affiliation":[{"name":"College of Intelligent Systems Science and Engineering, Hubei Minzu University, No. 39 Xueyuan Road, Enshi 445000, P.\u00a0R.\u00a0China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jun","family":"Li","sequence":"additional","affiliation":[{"name":"College of Intelligent Systems Science and Engineering, Hubei Minzu University, No. 39 Xueyuan Road, Enshi 445000, P.\u00a0R.\u00a0China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Tao","family":"Hu","sequence":"additional","affiliation":[{"name":"College of Intelligent Systems Science and Engineering, Hubei Minzu University, No. 39 Xueyuan Road, Enshi 445000, P.\u00a0R.\u00a0China"},{"name":"Hubei Engineering Research Center of Selenium Food Nutrition and Health Intelligent Technology, No. 39, Xueyuan Road, Enshi 445000, P. R. China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xuanyu","family":"Zhao","sequence":"additional","affiliation":[{"name":"College of Intelligent Systems Science and Engineering, Hubei Minzu University, No. 39 Xueyuan Road, Enshi 445000, P.\u00a0R.\u00a0China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"219","published-online":{"date-parts":[[2024,5,6]]},"reference":[{"key":"S0219467824500608BIB001","first-page":"1","volume-title":"31st Conf. Advances in Neural Information Processing Systems","volume":"30","author":"Vaswani A.","year":"2017"},{"key":"S0219467824500608BIB002","first-page":"4055","volume-title":"Proc. 35th Int. Conf. Machine Learning","author":"Parmar N.","year":"2018"},{"doi-asserted-by":"publisher","key":"S0219467824500608BIB003","DOI":"10.1007\/978-3-030-58452-8_13"},{"doi-asserted-by":"publisher","key":"S0219467824500608BIB005","DOI":"10.1109\/ICCV48922.2021.00062"},{"doi-asserted-by":"publisher","key":"S0219467824500608BIB006","DOI":"10.1109\/TPAMI.2022.3164083"},{"key":"S0219467824500608BIB007","first-page":"2286","volume-title":"Proc. 38th Int. Conf. Machine Learning","author":"D\u2019Ascoli S.","year":"2021"},{"doi-asserted-by":"publisher","key":"S0219467824500608BIB008","DOI":"10.1109\/CVPR52688.2022.01186"},{"doi-asserted-by":"publisher","key":"S0219467824500608BIB009","DOI":"10.1109\/CVPR46437.2021.01625"},{"key":"S0219467824500608BIB010","first-page":"3965","volume-title":"35th Conf. Advances in Neural Information Processing Systems","volume":"34","author":"Dai Z.","year":"2021"}],"container-title":["International Journal of Image and Graphics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.worldscientific.com\/doi\/pdf\/10.1142\/S0219467824500608","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,1,20]],"date-time":"2025-01-20T08:27:14Z","timestamp":1737361634000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.worldscientific.com\/doi\/10.1142\/S0219467824500608"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,5,6]]},"references-count":9,"journal-issue":{"issue":"01","published-print":{"date-parts":[[2025,1]]}},"alternative-id":["10.1142\/S0219467824500608"],"URL":"https:\/\/doi.org\/10.1142\/s0219467824500608","relation":{},"ISSN":["0219-4678","1793-6756"],"issn-type":[{"type":"print","value":"0219-4678"},{"type":"electronic","value":"1793-6756"}],"subject":[],"published":{"date-parts":[[2024,5,6]]},"article-number":"2450060"}}