{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,3]],"date-time":"2026-07-03T12:16:55Z","timestamp":1783081015351,"version":"3.54.6"},"reference-count":47,"publisher":"Oxford University Press (OUP)","issue":"5","license":[{"start":{"date-parts":[[2026,5,1]],"date-time":"2026-05-01T00:00:00Z","timestamp":1777593600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by-nc\/4.0\/"}],"funder":[{"DOI":"10.13039\/100019455","name":"University of Economics Ho Chi Minh City","doi-asserted-by":"publisher","id":[{"id":"10.13039\/100019455","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2026,5,7]]},"abstract":"<jats:title>Abstract<\/jats:title>\n                  <jats:p>Interior design style recognition presents a challenging computational problem due to fine-grained inter-style similarities, complex spatial arrangements, and strong variability in materials, lighting, and layout. From a computational design and engineering perspective, accurately modelling such stylistic variations is essential not only for classification, but also for enabling semantic representation, retrieval, and analysis of design cases within data-driven design systems. However, existing approaches often struggle to balance the extraction of fine-grained local cues with the modelling of global contextual relationships in complex interior scenes. In this study, we propose a unified data-model framework for interior design style recognition. First, we introduce InDes, a curated interior design image data set comprising 3826 images across five architectural styles. The data set integrates expert-labelled real-world samples from Vietnamese studios, with additional pseudo-labelled data generated through a semisupervised learning strategy, addressing class imbalance and supporting robust model training under limited annotation conditions. Second, we propose AT3 (Attention Transformer for Interior Style Design Classification), an attention-guided transformer architecture that combines convolutional feature extraction with hierarchical attention mechanisms and transformer-based global context modelling. By applying attention-based feature refinement prior to transformer tokenization, the proposed architecture preserves fine-grained stylistic cues while enabling effective long-range dependency modelling across interior scenes. This design improves the quality of learned representations for complex design environments characterized by subtle visual differences. Extensive experiments on the InDes data set demonstrate that AT3, when instantiated with an Xception backbone, achieves a validation accuracy of 87.8% and an F1-score of 0.86, outperforming strong convolutional and transformer-based baselines by margins of 5.2% and 6.7%, respectively. Ablation studies further validate the effectiveness of the proposed attention-guided architecture. Beyond classification performance, the learned representations provide a computational foundation for downstream design-oriented tasks such as style-aware retrieval, clustering, and data-driven analysis in computational design and engineering systems.<\/jats:p>","DOI":"10.1093\/jcde\/qwag044","type":"journal-article","created":{"date-parts":[[2026,5,1]],"date-time":"2026-05-01T12:07:16Z","timestamp":1777637236000},"page":"183-202","source":"Crossref","is-referenced-by-count":0,"title":["Attention-guided transformer architecture for fine-grained style representation in design systems"],"prefix":"10.1093","volume":"13","author":[{"ORCID":"https:\/\/orcid.org\/0000-0001-8341-7319","authenticated-orcid":false,"given":"Bao T","family":"Nguyen","sequence":"first","affiliation":[{"name":"Institute of Intelligent and Interactive Technologies, University of Economics Ho Chi Minh City ,","place":["Vietnam"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-0526-0736","authenticated-orcid":false,"given":"Thinh T","family":"Nguyen","sequence":"additional","affiliation":[{"name":"Institute of Intelligent and Interactive Technologies, University of Economics Ho Chi Minh City ,","place":["Vietnam"]}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"286","published-online":{"date-parts":[[2026,5,1]]},"reference":[{"key":"2026070307221063500_bib1","doi-asserted-by":"publisher","DOI":"10.4324\/9780429502668","volume-title":"A Philosophy of Interior Design","author":"Abercrombie","year":"2018"},{"key":"2026070307221063500_bib2","article-title":"Beit: Bert pre-training of image transformers","author":"Bao","year":"2021"},{"key":"2026070307221063500_bib3","doi-asserted-by":"publisher","first-page":"2601","DOI":"10.19139\/soic-2310-5070-2524","article-title":"Forecasting scientific impact: A model for predicting citation counts","volume":"13","author":"Bao","year":"2025","journal-title":"Statistics, Optimization & Information Computing"},{"key":"2026070307221063500_bib4","doi-asserted-by":"crossref","first-page":"3479","DOI":"10.1109\/CVPR.2015.7298970","article-title":"Material recognition in the wild with the materials in context database","volume-title":"2015 IEEE Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Bell","year":"2015"},{"key":"2026070307221063500_bib5","first-page":"138","article-title":"Interior Design: A Critical Introduction by Clive Edwards","volume-title":"Idea Journal","author":"Caan","year":"2011"},{"key":"2026070307221063500_bib6","doi-asserted-by":"publisher","first-page":"28125","DOI":"10.1007\/s11042-023-16579-0","article-title":"Hynet: A novel hybrid deep learning approach for efficient interior design texture retrieval","volume":"2023","author":"Chen","year":"2023","journal-title":"Multimedia Tools and Applications"},{"key":"2026070307221063500_bib7","volume-title":"Architecture: Form, Space, and Order","author":"Ching","year":"2014"},{"key":"2026070307221063500_bib8","doi-asserted-by":"publisher","first-page":"275","DOI":"10.1093\/jcde\/qwaf137","article-title":"Toward fully automated cad\u2013cae integration through design feature recognition and small language models","volume":"13","author":"Chun","year":"2025","journal-title":"Journal of Computational Design and Engineering"},{"key":"2026070307221063500_bib9","first-page":"3965","article-title":"Coatnet: marrying convolution and attention for all data sizes","volume":"34","author":"Dai","year":"2021","journal-title":"Advances in Neural Information Processing Systems (NeurIPS)"},{"key":"2026070307221063500_bib10","article-title":"An image is worth 16x16 words: Transformers for image recognition at scale","author":"Dosovitskiy","year":"2020"},{"key":"2026070307221063500_bib11","doi-asserted-by":"publisher","first-page":"1086","DOI":"10.1016\/j.rser.2017.02.011","article-title":"Parametric design and daylighting: A literature review","volume":"73","author":"Eltaweel","year":"2017","journal-title":"Renewable and Sustainable Energy Reviews"},{"key":"2026070307221063500_bib12","doi-asserted-by":"crossref","first-page":"87","DOI":"10.1109\/TPAMI.2022.3152247","article-title":"A survey on vision transformer","volume":"45","author":"Han","year":"2023","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"2026070307221063500_bib13","first-page":"16000","article-title":"Masked autoencoders are scalable vision learners","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR)","author":"He","year":"2022"},{"key":"2026070307221063500_bib14","doi-asserted-by":"publisher","first-page":"85","DOI":"10.1093\/jcde\/qwae017","article-title":"Generative artificial intelligence and building design: early photorealistic render visualization of fa\u00e7ades using local identity-trained models","volume":"11","author":"Jo","year":"2024","journal-title":"Journal of Computational Design and Engineering"},{"key":"2026070307221063500_bib15","volume-title":"Construction Drawings and Details for Interiors","author":"Kilmer","year":"2016"},{"key":"2026070307221063500_bib16","doi-asserted-by":"publisher","first-page":"7299","DOI":"10.3390\/app10207299","article-title":"Stochastic detection of interior design styles using a deep-learning model for reference images","volume":"10","author":"Kim","year":"2020","journal-title":"Applied Sciences"},{"key":"2026070307221063500_bib17","article-title":"Adam: A method for stochastic optimization","author":"Kingma","year":"2014"},{"key":"2026070307221063500_bib18","first-page":"1097","article-title":"Imagenet classification with deep convolutional neural networks","volume-title":"Advances in Neural Information Processing Systems","author":"Krizhevsky","year":"2012"},{"key":"2026070307221063500_bib19","doi-asserted-by":"publisher","first-page":"91","DOI":"10.1007\/s11768-024-00234-6","article-title":"Application of feedforward and recurrent neural networks for model-based control systems","volume":"23","author":"Krok","year":"2025","journal-title":"Control Theory and Technology"},{"key":"2026070307221063500_bib20","doi-asserted-by":"crossref","first-page":"436","DOI":"10.1038\/nature14539","article-title":"Deep learning","volume":"521","author":"LeCun","year":"2015","journal-title":"Nature"},{"key":"2026070307221063500_bib21","article-title":"Pseudo-label: The simple and efficient semi-supervised learning method for deep neural networks","author":"Lee","year":"2013"},{"key":"2026070307221063500_bib22","doi-asserted-by":"publisher","first-page":"8006069","DOI":"10.1155\/2022\/8006069","article-title":"Interior space design and automatic layout method based on cnn","volume":"2022","author":"Leung","year":"2022","journal-title":"Mathematical Problems in Engineering"},{"key":"2026070307221063500_bib23","doi-asserted-by":"publisher","first-page":"4","DOI":"10.1140\/epjds\/s13688-019-0182-z","article-title":"Inside 50,000 living rooms: an assessment of global residential ornamentation using transfer learning","volume":"8","author":"Liu","year":"2019","journal-title":"EPJ Data Science"},{"key":"2026070307221063500_bib24","first-page":"10012","article-title":"Hierarchical vision transformer using shifted windows","volume-title":"The IEEE\/CVF International Conference on Computer Vision (ICCV) 2021","author":"Liu","year":"2021"},{"key":"2026070307221063500_bib25","article-title":"Sgdr: Stochastic gradient descent with warm restarts","author":"Loshchilov","year":"2017"},{"key":"2026070307221063500_bib26","first-page":"100","volume-title":"Interior Design","author":"Pile","year":"2003"},{"key":"2026070307221063500_bib27","volume-title":"Professional Practice for Interior Designers","author":"Piotrowski","year":"2020"},{"key":"2026070307221063500_bib28","doi-asserted-by":"crossref","DOI":"10.5040\/9781501371547","volume-title":"Meanings of Designed Spaces","author":"Poldma","year":"2013"},{"key":"2026070307221063500_bib29","first-page":"8748","article-title":"Learning transferable visual models from natural language supervision","volume-title":"International Conference on Machine Learning (ICML)","author":"Radford","year":"2021"},{"key":"2026070307221063500_bib30","first-page":"8821","article-title":"Zero-shot text-to-image generation. International Conference on Machine Learning","author":"Ramesh","year":"2021"},{"key":"2026070307221063500_bib31","doi-asserted-by":"publisher","first-page":"157","DOI":"10.1093\/jcde\/qwag029","article-title":"Learning design preferences through design feature extraction and weighted ensemble","volume":"13","author":"Shin","year":"2026","journal-title":"Journal of Computational Design and Engineering"},{"key":"2026070307221063500_bib32","first-page":"596","article-title":"Fixmatch: Simplifying semi-supervised learning with consistency and confidence","volume-title":"The 34th International Conference on Neural Information Processing Systems (NIPS '20)","author":"Sohn","year":"2020"},{"key":"2026070307221063500_bib33","first-page":"1929","article-title":"Dropout: A simple way to prevent neural networks from overfitting","volume":"15","author":"Srivastava","year":"2014","journal-title":"Journal of Machine Learning Research"},{"key":"2026070307221063500_bib34","doi-asserted-by":"publisher","first-page":"103787","DOI":"10.1016\/j.cities.2022.103787","article-title":"Understanding architecture age and style through deep learning","volume":"128","author":"Sun","year":"2022","journal-title":"Cities"},{"key":"2026070307221063500_bib35","doi-asserted-by":"publisher","first-page":"92","DOI":"10.1177\/10717641251324234","article-title":"Developing visual literacy in the first-year interior design studio","volume":"50","author":"Tooley","year":"2025","journal-title":"Journal of Interior Design"},{"key":"2026070307221063500_bib36","first-page":"10347","article-title":"Training data-efficient image transformers and distillation through attention","volume":"139","author":"Touvron","year":"2021","journal-title":"Proceedings of the International Conference on Machine Learning (ICML)"},{"key":"2026070307221063500_bib37","doi-asserted-by":"crossref","first-page":"739","DOI":"10.1109\/GTSD.2018.8595551","article-title":"Facial expression recognition based on salient regions","volume-title":"2018 4th International Conference on Green Technology and Sustainable Development (GTSD)","author":"Vo","year":"2018"},{"key":"2026070307221063500_bib38","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1007\/s00521-023-08431-1","article-title":"An efficient framework for outfit compatibility prediction towards occasion","volume":"35","author":"Vo","year":"2023","journal-title":"Neural Computing and Applications"},{"key":"2026070307221063500_bib39","doi-asserted-by":"publisher","first-page":"126972","DOI":"10.1016\/j.neucom.2023.126972","article-title":"A framework-based transformer and knowledge distillation for interior style classification","volume":"565","author":"Vo","year":"2024","journal-title":"Neurocomputing"},{"key":"2026070307221063500_bib40","doi-asserted-by":"publisher","first-page":"3","DOI":"10.1007\/978-3-030-01234-2_1","article-title":"Cbam: Convolutional block attention module","volume-title":"In Computer Vision \u2013 ECCV 2018: 15th European Conference, Munich, Germany, September 8\u201314, 2018, Proceedings, Part VII","author":"Woo","year":"2018"},{"key":"2026070307221063500_bib41","article-title":"Demystify self-attention in vision transformers from a semantic perspective: Analysis and application","author":"Wu","year":"2022"},{"key":"2026070307221063500_bib42","doi-asserted-by":"publisher","first-page":"10684","DOI":"10.1109\/CVPR42600.2020.01070","volume-title":"Self-training with Noisy Student Improves Imagenet Classification.","author":"Xie","year":"2020"},{"key":"2026070307221063500_bib43","doi-asserted-by":"publisher","DOI":"10.3390\/rs15092406","article-title":"A cnn-transformer network combining cbam for change detection in high-resolution remote sensing images","volume":"15","author":"Yin","year":"2023","journal-title":"Remote Sensing"},{"key":"2026070307221063500_bib44","doi-asserted-by":"publisher","first-page":"23729","DOI":"10.1007\/s11042-020-08976-6","article-title":"A review of object detection based on deep learning","volume":"79","author":"Youzi","year":"2020","journal-title":"Multimedia Tools and Applications"},{"key":"2026070307221063500_bib45","article-title":"Indoor space recognition using deep convolutional neural network: A case study at mit campus","author":"Zhang","year":"2016"},{"key":"2026070307221063500_bib46","article-title":"Generalized cross entropy loss for training deep neural networks with noisy labels","author":"Zhang","year":"2018"},{"key":"2026070307221063500_bib47","doi-asserted-by":"publisher","first-page":"1452","DOI":"10.1109\/TPAMI.2017.2723009","article-title":"Places: A 10 million image database for scene recognition","volume":"40","author":"Zhou","year":"2018","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"}],"container-title":["Journal of Computational Design and Engineering"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/academic.oup.com\/jcde\/advance-article-pdf\/doi\/10.1093\/jcde\/qwag044\/68201474\/qwag044.pdf","content-type":"application\/pdf","content-version":"am","intended-application":"syndication"},{"URL":"https:\/\/academic.oup.com\/jcde\/article-pdf\/13\/5\/183\/68201474\/qwag044.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/academic.oup.com\/jcde\/article-pdf\/13\/5\/183\/68201474\/qwag044.pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,7,3]],"date-time":"2026-07-03T11:22:22Z","timestamp":1783077742000},"score":1,"resource":{"primary":{"URL":"https:\/\/academic.oup.com\/jcde\/article\/13\/5\/183\/8666383"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,5,1]]},"references-count":47,"journal-issue":{"issue":"5","published-print":{"date-parts":[[2026,5,7]]}},"URL":"https:\/\/doi.org\/10.1093\/jcde\/qwag044","relation":{},"ISSN":["2288-5048"],"issn-type":[{"value":"2288-5048","type":"electronic"}],"subject":[],"published-other":{"date-parts":[[2026,5]]},"published":{"date-parts":[[2026,5,1]]}}}