{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,31]],"date-time":"2026-07-31T08:57:57Z","timestamp":1785488277850,"version":"3.56.0"},"publisher-location":"New York, NY, USA","reference-count":17,"publisher":"ACM","license":[{"start":{"date-parts":[[2025,12,17]],"date-time":"2025-12-17T00:00:00Z","timestamp":1765929600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/legalcode"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2025,12,17]]},"DOI":"10.1145\/3774521.3774581","type":"proceedings-article","created":{"date-parts":[[2026,7,31]],"date-time":"2026-07-31T07:34:24Z","timestamp":1785483264000},"page":"1-9","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["Strongly Connected Components Are All You Need: Graph-Theoretic Interpretability and Optimization for Vision Transformers"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0009-0000-8210-2526","authenticated-orcid":false,"given":"Devansh","family":"Garg","sequence":"first","affiliation":[{"name":"Indian Institute of Technology Mandi, Mandi, Himachal Pradesh, India"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2026,7,31]]},"reference":[{"key":"e_1_3_3_1_2_2","unstructured":"Alexey Dosovitskiy Lucas Beyer Alexander Kolesnikov Dirk Weissenborn Xiaohua Zhai Thomas Unterthiner Mostafa Dehghani Matthias Minderer Georg Heigold Sylvain Gelly et\u00a0al. An image is worth 16x16 words: Transformers for image recognition at scale. arXiv:https:\/\/arXiv.org\/abs\/2010.11929. 2020. Retrieved from https:\/\/arxiv.org\/abs\/2010.11929"},{"key":"e_1_3_3_1_3_2","unstructured":"Karen Simonyan and Andrew Zisserman. Very deep convolutional networks for large-scale image recognition. arXiv:https:\/\/arXiv.org\/abs\/1409.1556. 2014. Retrieved from https:\/\/arxiv.org\/abs\/1409.1556"},{"key":"e_1_3_3_1_4_2","doi-asserted-by":"publisher","unstructured":"Sarah Wiegreffe and Yuval Pinter. Attention is not not explanation. In Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP \u201919). Association for Computational Linguistics Hong Kong China 11\u201320. 2019. 10.18653\/v1\/D19-1002","DOI":"10.18653\/v1\/D19-1002"},{"key":"e_1_3_3_1_5_2","unstructured":"Mukund Sundararajan Ankur Taly and Qiqi Yan. Axiomatic attribution for deep networks. In Proceedings of the 34th International Conference on Machine Learning (ICML \u201917) PMLR 70 3319\u20133328. 2017."},{"key":"e_1_3_3_1_6_2","doi-asserted-by":"publisher","unstructured":"Robert E. Tarjan. Depth-first search and linear graph algorithms. SIAM J. Comput. 1 2 (June 1972) 146\u2013160. 10.1137\/0201010","DOI":"10.1137\/0201010"},{"key":"e_1_3_3_1_7_2","unstructured":"Ashish Vaswani Noam Shazeer Niki Parmar Jakob Uszkoreit Llion Jones Aidan N. Gomez \u0141ukasz Kaiser and Illia Polosukhin. Attention is all you need. In Proceedings of the 31st International Conference on Neural Information Processing Systems (NIPS \u201917). Curran Associates Inc. Red Hook NY USA 6000\u20136010. 2017."},{"key":"e_1_3_3_1_8_2","unstructured":"Hugo Touvron Matthieu Cord Matthijs Douze Francisco Massa Alexandre Sablayrolles and Herv\u00e9 J\u00e9gou. Training data-efficient image transformers & distillation through attention. In Proceedings of the 38th International Conference on Machine Learning (ICML \u201921) PMLR 139 10347\u201310357. 2021."},{"key":"e_1_3_3_1_9_2","unstructured":"Alec Radford Jong Wook Kim Chris Hallacy Aditya Ramesh Gabriel Goh Sandhini Agarwal Girish Sastry Amanda Askell Pamela Mishkin Jack Clark et\u00a0al. Learning transferable visual representations from natural language supervision. In Proceedings of the 38th International Conference on Machine Learning (ICML \u201921) PMLR 139 8748\u20138763. 2021."},{"key":"e_1_3_3_1_10_2","doi-asserted-by":"publisher","unstructured":"Zonghan Wu Shirui Pan Fengwen Chen Guodong Long Chengqi Zhang and Philip S. Yu. A comprehensive survey on graph neural networks. IEEE Trans. Neural Netw. Learn. Syst. 32 1 (Jan. 2021) 4\u201324. 10.1109\/TNNLS.2020.2978386","DOI":"10.1109\/TNNLS.2020.2978386"},{"key":"e_1_3_3_1_11_2","doi-asserted-by":"publisher","unstructured":"Yann LeCun Yoshua Bengio and Geoffrey Hinton. Deep learning. Nature 521 7553 (2015) 436\u2013444. 10.1038\/nature14539","DOI":"10.1038\/nature14539"},{"key":"e_1_3_3_1_12_2","unstructured":"Ian Goodfellow Yoshua Bengio and Aaron Courville. Deep Learning. MIT Press Cambridge MA USA. 2016."},{"key":"e_1_3_3_1_13_2","doi-asserted-by":"publisher","unstructured":"Kaiming He Xiangyu Zhang Shaoqing Ren and Jian Sun. Deep residual learning for image recognition. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR \u201916) Las Vegas NV USA. IEEE 770\u2013778. 2016. 10.1109\/CVPR.2016.90","DOI":"10.1109\/CVPR.2016.90"},{"key":"e_1_3_3_1_14_2","doi-asserted-by":"publisher","unstructured":"Jacob Devlin Ming-Wei Chang Kenton Lee and Kristina Toutanova. BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT \u201919). Association for Computational Linguistics Minneapolis Minnesota 4171\u20134186. 2019. 10.18653\/v1\/N19-1423","DOI":"10.18653\/v1\/N19-1423"},{"key":"e_1_3_3_1_15_2","doi-asserted-by":"publisher","unstructured":"Hila Chefer Shir Gur and Lior Wolf. Transformer interpretability beyond attention visualization. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR \u201921) Nashville TN USA. IEEE 782\u2013791. 2021. 10.1109\/CVPR46437.2021.00084","DOI":"10.1109\/CVPR46437.2021.00084"},{"key":"e_1_3_3_1_16_2","unstructured":"Samira Abnar and Willem Zuidema. Quantifying attention flow in transformers. arXiv:https:\/\/arXiv.org\/abs\/2005.00928. 2020. Retrieved from https:\/\/arxiv.org\/abs\/2005.00928"},{"key":"e_1_3_3_1_17_2","unstructured":"Song Han Jeff Pool John Tran and William J. Dally. Learning both weights and connections for efficient neural network. In Proceedings of the 28th International Conference on Neural Information Processing Systems (NIPS \u201915). MIT Press Cambridge MA USA 1135\u20131143. 2015."},{"key":"e_1_3_3_1_18_2","unstructured":"Andreas Steiner Alexander Kolesnikov Xiaohua Zhai Ross Wightman Jakob Uszkoreit and Lucas Beyer. How to train your ViT? Data augmentation and regularization in Vision Transformers. arXiv:https:\/\/arXiv.org\/abs\/2106.10270. 2021. Retrieved from https:\/\/arxiv.org\/abs\/2106.10270"}],"event":{"name":"ICVGIP 2025: Indian Conference on Computer Vision, Graphics, and Image Processing","location":"Mandi Himachal Pradesh India","acronym":"ICVGIP 2025"},"container-title":["Proceedings of the Sixteen Indian Conference on Computer Vision, Graphics and Image Processing"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3774521.3774581","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,7,31]],"date-time":"2026-07-31T08:05:12Z","timestamp":1785485112000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3774521.3774581"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,12,17]]},"references-count":17,"alternative-id":["10.1145\/3774521.3774581","10.1145\/3774521"],"URL":"https:\/\/doi.org\/10.1145\/3774521.3774581","relation":{},"subject":[],"published":{"date-parts":[[2025,12,17]]},"assertion":[{"value":"2026-07-31","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}