{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,8,18]],"date-time":"2026-08-18T07:29:21Z","timestamp":1787038161752,"version":"build-2736575974"},"reference-count":34,"publisher":"Springer Science and Business Media LLC","issue":"8","license":[{"start":{"date-parts":[[2026,7,28]],"date-time":"2026-07-28T00:00:00Z","timestamp":1785196800000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/www.springernature.com\/gp\/researchers\/text-and-data-mining"},{"start":{"date-parts":[[2026,7,28]],"date-time":"2026-07-28T00:00:00Z","timestamp":1785196800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.springernature.com\/gp\/researchers\/text-and-data-mining"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Computing"],"published-print":{"date-parts":[[2026,8]]},"DOI":"10.1007\/s00607-026-01711-3","type":"journal-article","created":{"date-parts":[[2026,7,28]],"date-time":"2026-07-28T09:41:53Z","timestamp":1785231713000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":0,"title":["Structural semantic organization distillation for cross-task transfer from prompt-conditioned foundation models to CNN detectors"],"prefix":"10.1007","volume":"108","author":[{"given":"Fucheng","family":"Li","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yongan","family":"Guo","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Guodong","family":"Wang","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2026,7,28]]},"reference":[{"key":"1711_CR1","doi-asserted-by":"publisher","unstructured":"Kirillov A, Mintun E, Ravi N, Mao H, Rolland C, Gustafson L, Xiao T, Whitehead S, Berg AC, Lo W-Y, Doll\u00e1r P, Girshick R (2023) Segment anything. In: Proceedings of the IEEE\/CVF international conference on computer vision (ICCV), pp. 4015\u20134026 . https:\/\/doi.org\/10.1109\/ICCV51070.2023.00371","DOI":"10.1109\/ICCV51070.2023.00371"},{"key":"1711_CR2","unstructured":"Ravi N, Gabeur V, Hu Y-T, Hu R, Ryali C, Ma T, Khedr H, R\u00e4dle R, Rolland C, Gustafson L, Mintun E, Pan J, Alwala KV, Carion N, Wu C-Y, Girshick R, Doll\u00e1r P, Feichtenhofer C (2024) SAM 2: segment anything in images and videos. arxiv: 2408.00714"},{"key":"1711_CR3","unstructured":"Carion N, Gustafson L, Hu Y-T, Debnath S, Hu R, Suris D, Ryali C, Alwala KV, Khedr H, et al (2025) SAM 3: segment anything with concepts. arxiv: 2511.16719"},{"key":"1711_CR4","doi-asserted-by":"publisher","unstructured":"Radford A, Kim JW, Hallacy C, Ramesh A, Goh G, Agarwal S, Sastry G, Askell A, Mishkin P, Clark J, Krueger G, Sutskever I (2021) Learning transferable visual models from natural language supervision. In: Proceedings of the 38th international conference on machine learning (ICML), pp. 8748\u20138763. https:\/\/doi.org\/10.48550\/arXiv.2103.00020","DOI":"10.48550\/arXiv.2103.00020"},{"key":"1711_CR5","doi-asserted-by":"publisher","unstructured":"Liu S, Zeng Z, Ren T, Li F, Zhang H, Yang J, Li C, Yang J, Su H, Zhu J, Zhang L (2024) Grounding DINO: marrying DINO with grounded pre-training for open-set object detection. In: Proceedings of the European conference on computer vision (ECCV). https:\/\/doi.org\/10.48550\/arXiv.2303.05499","DOI":"10.48550\/arXiv.2303.05499"},{"key":"1711_CR6","doi-asserted-by":"publisher","unstructured":"Minderer M, Gritsenko A, Stone A, Neumann M, Weissenborn D, Dosovitskiy A, Mahendran A, Arnab A, Dehghani M, Shen Z, Wang X, Zhai X, Kipf T, Houlsby N (2022) Simple open-vocabulary object detection with vision transformers. In: Proceedings of the European conference on computer vision (ECCV), pp. 728\u2013755. https:\/\/doi.org\/10.1007\/978-3-031-20083-0_42","DOI":"10.1007\/978-3-031-20083-0_42"},{"key":"1711_CR7","unstructured":"Yang Z, Li Z, Zeng A, Li Z, Yuan C, Li Y (2024) ViTKD: Practical guidelines for ViT feature knowledge distillation. In: Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition workshops (CVPRW). arxiv: 2209.02432"},{"key":"1711_CR8","doi-asserted-by":"crossref","unstructured":"Fan J, Li C, Liu X, Yao A (2024) ScaleKD: Strong vision transformers could be excellent teachers. In: Advances in neural information processing systems (NeurIPS). arxiv: 2411.06786","DOI":"10.52202\/079017-2022"},{"key":"1711_CR9","doi-asserted-by":"crossref","unstructured":"Hao Z, Guo J, Han K, Tang Y, Hu H, Wang Y, Xu C One-for-all: Bridge the gap between heterogeneous architectures in knowledge distillation. Advances in neural information processing systems (NeurIPS) (2023) arxiv: 2310.19444","DOI":"10.52202\/075280-3483"},{"key":"1711_CR10","doi-asserted-by":"crossref","unstructured":"Wang J, Chen Y, Zheng Z, Li X, Cheng M-M, Hou Q (2024) CrossKD: Cross-head knowledge distillation for object detection. In: Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition (CVPR), pp. 16520\u201316530. arxiv: 2306.11369","DOI":"10.1109\/CVPR52733.2024.01563"},{"key":"1711_CR11","unstructured":"Hinton G, Vinyals O, Dean J (2015) Distilling the knowledge in a neural network. NIPS 2014 deep learning workshop. arxiv: 1503.02531"},{"key":"1711_CR12","unstructured":"Zheng Z, Ye R, Wang P, Ren D, Zuo W, Hou Q, Cheng M-M (2022) Localization distillation for dense object detection. In: Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition (CVPR). pp 9407\u20139416 arxiv: 2102.12252"},{"key":"1711_CR13","unstructured":"Yang Z, Li Z, Jiang X, Gong Y, Yuan Z, Zhao D, Yuan C (2022) Focal and global knowledge distillation for detectors. In: Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition (CVPR), pp. 4643\u20134652. arxiv: 2111.11837"},{"key":"1711_CR14","doi-asserted-by":"publisher","unstructured":"Yang Z, Li Z, Shao M, Shi D, Yuan Z, Yuan C (2022) Masked generative distillation. In: Proceedings of the European conference on computer vision (ECCV), pp. 53\u201369. https:\/\/doi.org\/10.1007\/978-3-031-20083-0_4","DOI":"10.1007\/978-3-031-20083-0_4"},{"key":"1711_CR15","doi-asserted-by":"crossref","unstructured":"Cao W, Zhang Y, Gao J, Cheng A, Cheng K, Cheng J (2022) PKD: general distillation framework for object detectors via pearson correlation coefficient. In: Advances in neural information processing systems (NeurIPS) . arxiv: 2207.02039","DOI":"10.52202\/068431-1120"},{"key":"1711_CR16","unstructured":"Du Z, Zhang R, Chang M, Zhang X, Liu S, Chen T, Chen Y (2021) Distilling object detectors with feature richness. In: Advances in neural information processing systems (NeurIPS). arxiv: 2111.00674"},{"key":"1711_CR17","doi-asserted-by":"crossref","unstructured":"Zhang Z, Li J, Li J, Xu J (2025) SAMKD: spatial-aware adaptive masking knowledge distillation for object detection. arxiv: 2501.07101","DOI":"10.2139\/ssrn.5002082"},{"key":"1711_CR18","doi-asserted-by":"crossref","unstructured":"Lin J-H, Yao Y, Hsu C-F, Xie H-X, Shuai H-H, Cheng W-H (2025) Perspective-aware teaching: adapting knowledge for heterogeneous distillation. arxiv: 2501.08885","DOI":"10.1109\/ICCV51701.2025.00398"},{"key":"1711_CR19","unstructured":"Peng Y, Ye H, Huang C, Hu X, Chen J, Zeng R (2025) Revisiting cross-architecture distillation: adaptive dual-teacher transfer for lightweight video models. arxiv: 2511.09469"},{"key":"1711_CR20","unstructured":"Park W, Kim D, Lu Y, Cho M (2019) Relational knowledge distillation. In: Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition (CVPR). pp 3967\u20133976 arxiv: 1904.05068"},{"key":"1711_CR21","doi-asserted-by":"publisher","unstructured":"Liu Y, Cao J, Li B, Yuan C, Hu W, Li Y, Duan Y (2019) Knowledge distillation via instance relationship graph. In: Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition (CVPR), pp. 7096\u20137104. https:\/\/doi.org\/10.1109\/CVPR.2019.00727","DOI":"10.1109\/CVPR.2019.00727"},{"key":"1711_CR22","doi-asserted-by":"crossref","unstructured":"Liu L, Huang Q, Lin S, Xie H, Wang B, Chang X, Liang X (2021) Exploring inter-channel correlation for diversity-preserved knowledge distillation. In: Proceedings of the IEEE\/CVF international conference on computer vision (ICCV). pp 8271\u20138280","DOI":"10.1109\/ICCV48922.2021.00816"},{"key":"1711_CR23","doi-asserted-by":"crossref","unstructured":"Tung F, Mori G (2019) Similarity-preserving knowledge distillation. In: Proceedings of the IEEE\/CVF international conference on computer vision (ICCV). pp 1365\u20131374 arxiv: 1907.09682","DOI":"10.1109\/ICCV.2019.00145"},{"key":"1711_CR24","doi-asserted-by":"crossref","unstructured":"Huang T, You S, Wang F, Qian C, Xu C Knowledge distillation from a stronger teacher. Advances in neural information processing systems (NeurIPS) (2022) arxiv: 2205.10536","DOI":"10.52202\/068431-2443"},{"key":"1711_CR25","doi-asserted-by":"crossref","unstructured":"Zhang W, Xie F, Cai W, Ma C VRM: Knowledge distillation via virtual relation matching. In: Proceedings of the IEEE\/CVF international conference on computer vision (ICCV) (2025) arxiv: 2502.20760","DOI":"10.1109\/ICCV51701.2025.00260"},{"key":"1711_CR26","unstructured":"Zhang S, Liu H, He K (2023) Knowledge distillation via token-level relationship graph. arxiv: 2306.12442"},{"key":"1711_CR27","unstructured":"Oquab M, Darcet T, Moutakanni T, Vo H, Szafraniec M, Khalidov V, Fernandez P, Haziza D, Massa F, El-Nouby A (2024) DINOv2: learning robust visual features without supervision. Transactions on machine learning research. arXiv:2304.07193"},{"key":"1711_CR28","doi-asserted-by":"crossref","unstructured":"Shu H, Li W, Tang Y, Zhang Y, Chen Y, Li H, Wang Y, Chen X (2025) TinySAM: pushing the envelope for efficient segment anything model. In: Proceedings of the AAAI conference on artificial intelligence (AAAI). arxiv: 2312.13789","DOI":"10.1609\/aaai.v39i19.34255"},{"key":"1711_CR29","doi-asserted-by":"publisher","unstructured":"Zhou C, Li X, Loy CC, Dai B (2023) EdgeSAM: prompt-in-the-loop distillation for on-device deployment of SAM. https:\/\/doi.org\/10.1007\/s11263-025-02562-9","DOI":"10.1007\/s11263-025-02562-9"},{"key":"1711_CR30","unstructured":"Zeng C, Jiang Y, Zhang A (2025) EfficientSAM3: progressive hierarchical distillation for video concept segmentation from SAM1, 2. p 3 arxiv: 2511.15833"},{"key":"1711_CR31","unstructured":"Chen T, Mai Z, Li R, Chao W-L (2023) Segment anything model (SAM) enhanced pseudo labels for weakly supervised semantic segmentation. arxiv: 2305.05803"},{"key":"1711_CR32","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2025.3554033","author":"J Wu","year":"2023","unstructured":"Wu J, Xu R, Wood-Doughty Z, Wang C, Xu S, Lam EY (2023) Segment anything model is a good teacher for local feature. Learning. https:\/\/doi.org\/10.1109\/TIP.2025.3554033","journal-title":"Learning"},{"key":"1711_CR33","doi-asserted-by":"publisher","unstructured":"Hu X, Sun F, Liu J, Xu F, Zhang X (2025) ST-SAM: SAM-driven self-training framework for semi-supervised camouflaged object detection. In: Proceedings of the 33rd ACM international conference on multimedia (MM). https:\/\/doi.org\/10.1145\/3746027.3755355","DOI":"10.1145\/3746027.3755355"},{"key":"1711_CR34","unstructured":"Ceausescu C-M, Anghelina I-M, Alexe D-B (2026) Multi-dataset cross-domain knowledge distillation for unified medical image segmentation, classification, and detection. arxiv: 2605.01563"}],"container-title":["Computing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s00607-026-01711-3.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s00607-026-01711-3","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s00607-026-01711-3.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,8,18]],"date-time":"2026-08-18T07:20:29Z","timestamp":1787037629000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s00607-026-01711-3"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,7,28]]},"references-count":34,"journal-issue":{"issue":"8","published-print":{"date-parts":[[2026,8]]}},"alternative-id":["1711"],"URL":"https:\/\/doi.org\/10.1007\/s00607-026-01711-3","relation":{},"ISSN":["0010-485X","1436-5057"],"issn-type":[{"value":"0010-485X","type":"print"},{"value":"1436-5057","type":"electronic"}],"subject":[],"published":{"date-parts":[[2026,7,28]]},"assertion":[{"value":"1 June 2026","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"30 June 2026","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"28 July 2026","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"The authors declare no conflict of interest.","order":1,"name":"Ethics","label":"Conflict of interest","group":{"name":"EthicsHeading","label":"Declarations"}}],"article-number":"127"}}