{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,21]],"date-time":"2026-07-21T23:04:02Z","timestamp":1784675042561,"version":"3.55.0"},"reference-count":66,"publisher":"Association for Computing Machinery (ACM)","issue":"1","license":[{"start":{"date-parts":[[2022,1,27]],"date-time":"2022-01-27T00:00:00Z","timestamp":1643241600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Multimedia Comput. Commun. Appl."],"published-print":{"date-parts":[[2022,1,31]]},"abstract":"<jats:p>Logo detection has been gaining considerable attention because of its wide range of applications in the multimedia field, such as copyright infringement detection, brand visibility monitoring, and product brand management on social media. In this article, we introduce LogoDet-3K, the largest logo detection dataset with full annotation, which has 3,000 logo categories, about 200,000 manually annotated logo objects, and 158,652 images. LogoDet-3K creates a more challenging benchmark for logo detection, for its higher comprehensive coverage and wider variety in both logo categories and annotated objects compared with existing datasets. We describe the collection and annotation process of our dataset and analyze its scale and diversity in comparison to other datasets for logo detection. We further propose a strong baseline method Logo-Yolo, which incorporates Focal loss and CIoU loss into the basic YOLOv3 framework for large-scale logo detection. It obtains about 4% improvement on the average performance compared with YOLOv3, and greater improvements compared with reported several deep detection models on LogoDet-3K. We perform extensive evaluation on three other existing datasets to further verify on both logo detection and retrieval tasks, and we demonstrate better generalization ability of LogoDet-3K on logo detection and retrieval tasks. The LogoDet-3K dataset is used to promote large-scale logo-related research. The code and LogoDet-3K can be found at https:\/\/github.com\/Wangjing1551\/LogoDet-3K-Dataset.<\/jats:p>","DOI":"10.1145\/3466780","type":"journal-article","created":{"date-parts":[[2022,1,27]],"date-time":"2022-01-27T19:44:21Z","timestamp":1643312661000},"page":"1-19","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":55,"title":["LogoDet-3K: A Large-scale Image Dataset for Logo Detection"],"prefix":"10.1145","volume":"18","author":[{"given":"Jing","family":"Wang","sequence":"first","affiliation":[{"name":"School of Information Science and Engineering, Shandong Normal University, Shandong, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Weiqing","family":"Min","sequence":"additional","affiliation":[{"name":"Key Laboratory of Intelligent Information Processing, Institute of Computing Technology, Chinese Academy of Sciences, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-4244-9207","authenticated-orcid":false,"given":"Sujuan","family":"Hou","sequence":"additional","affiliation":[{"name":"School of Information Science andEngineering, Shandong Normal University, Shandong, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Shengnan","family":"Ma","sequence":"additional","affiliation":[{"name":"School of Information Science and Engineering, Shandong Normal University, Shandong, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yuanjie","family":"Zheng","sequence":"additional","affiliation":[{"name":"School of Information Science and Engineering, Shandong Normal University, Shandong, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Shuqiang","family":"Jiang","sequence":"additional","affiliation":[{"name":"Key Laboratory of Intelligent Information Processing, Institute of Computing Technology, Chinese Academy of Sciences, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2022,1,27]]},"reference":[{"key":"e_1_3_2_2_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-23234-8_41"},{"key":"e_1_3_2_3_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.neucom.2017.03.051"},{"key":"e_1_3_2_4_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00644"},{"key":"e_1_3_2_5_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58452-8_13"},{"key":"e_1_3_2_6_2","doi-asserted-by":"publisher","DOI":"10.1145\/2964284.2964326"},{"key":"e_1_3_2_7_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2009.5206848"},{"key":"e_1_3_2_8_2","doi-asserted-by":"publisher","DOI":"10.1145\/3078971.3078990"},{"key":"e_1_3_2_9_2","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-009-0275-4"},{"key":"e_1_3_2_10_2","doi-asserted-by":"publisher","DOI":"10.1109\/WACV.2019.00081"},{"key":"e_1_3_2_11_2","article-title":"LOGO-Net: Large-scale deep logo detection and brand recognition with deep region-based convolutional networks","author":"Hoi Steven C. H.","year":"2015","unstructured":"Steven C. H. Hoi, Xiongwei Wu, Hantang Liu, Yue Wu, Huiqiong Wang, Hui Xue, and Qiang Wu. 2015. LOGO-Net: Large-scale deep logo detection and brand recognition with deep region-based convolutional networks. Retrieved from https:\/\/arXiv:1511.02462.","journal-title":"Retrieved from https:\/\/arXiv:1511.02462"},{"key":"e_1_3_2_12_2","article-title":"DeepLogo: Hitting logo recognition with the deep neural network hammer","author":"Iandola Forrest N.","year":"2015","unstructured":"Forrest N. Iandola, Anting Shen, Peter Gao, and Kurt Keutzer. 2015. DeepLogo: Hitting logo recognition with the deep neural network hammer. Retrieved from https:\/\/arXiv:1510.02131.","journal-title":"Retrieved from https:\/\/arXiv:1510.02131"},{"key":"e_1_3_2_13_2","doi-asserted-by":"publisher","DOI":"10.5555\/3157096.3157139"},{"key":"e_1_3_2_14_2","first-page":"346","volume-title":"Proceedings of the IEEE International Conference on Computer Vision","author":"Zhang Shaoqing Ren, Jian Sun, Kaiming He, and Xiangyu","year":"2014","unstructured":"Shaoqing Ren, Jian Sun, Kaiming He, and Xiangyu Zhang. 2014. Spatial pyramid pooling in deep convolutional networks for visual recognition. In Proceedings of the IEEE International Conference on Computer Vision. 346\u2013361."},{"key":"e_1_3_2_15_2","doi-asserted-by":"publisher","DOI":"10.1145\/1991996.1992016"},{"key":"e_1_3_2_16_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2011.6126456"},{"key":"e_1_3_2_17_2","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2020.3002345"},{"key":"e_1_3_2_18_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-59006-2_8"},{"key":"e_1_3_2_19_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01264-9_45"},{"key":"e_1_3_2_20_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.519"},{"key":"e_1_3_2_21_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.106"},{"key":"e_1_3_2_22_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.324"},{"key":"e_1_3_2_23_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-10602-1_48"},{"key":"e_1_3_2_24_2","first-page":"31","volume-title":"Computer Vision Image Understanding","author":"Tian Wengang Zhou, Bo Zhang, Lingxi Xie, and Qi","year":"2014","unstructured":"Wengang Zhou, Bo Zhang, Lingxi Xie, and Qi Tian. 2014. Fast and accurate near-duplicate image search with affinity propagation on the ImageWeb. In Computer Vision Image Understanding. Elsevier, 31\u201341."},{"key":"e_1_3_2_25_2","doi-asserted-by":"publisher","DOI":"10.1145\/3365212"},{"key":"e_1_3_2_26_2","first-page":"71","volume-title":"Proceedings of the AAAI Conference on Artificial Intelligence","author":"Liu Liu","year":"2018","unstructured":"Liu Liu, Daria Dzyabura, and Natalie Mizik. 2018. Visual listening in: Extracting brand image portrayed on social media. In Proceedings of the AAAI Conference on Artificial Intelligence. 71\u201377."},{"key":"e_1_3_2_27_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-46448-0_2"},{"key":"e_1_3_2_28_2","doi-asserted-by":"publisher","DOI":"10.5555\/850924.851523"},{"key":"e_1_3_2_29_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-10590-1_53"},{"key":"e_1_3_2_30_2","doi-asserted-by":"publisher","DOI":"10.1145\/1291233.1291467"},{"key":"e_1_3_2_31_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2005.177"},{"key":"e_1_3_2_32_2","doi-asserted-by":"publisher","DOI":"10.5555\/645651.665173"},{"key":"e_1_3_2_33_2","doi-asserted-by":"publisher","DOI":"10.1109\/IJCNN.2016.7727305"},{"key":"e_1_3_2_34_2","first-page":"2241","volume-title":"Proceedings of the Conference on Computer Vision and Pattern Recognition","author":"Girshick David A. McAllester Pedro F. Felzenszwalb, and Ross B.","year":"2010","unstructured":"David A. McAllester Pedro F. Felzenszwalb, and Ross B. Girshick. 2010. Cascade object detection with deformable part models. In Proceedings of the Conference on Computer Vision and Pattern Recognition. 2241\u20132248."},{"key":"e_1_3_2_35_2","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2009.167"},{"key":"e_1_3_2_36_2","first-page":"1","volume-title":"Proceedings of the Conference on Computer Vision and Pattern Recognition","author":"McAllester Deva Ramanan, Pedro F. Felzenszwalb, and David A.","year":"2008","unstructured":"Deva Ramanan, Pedro F. Felzenszwalb, and David A. McAllester. 2008. A discriminatively trained, multiscale, deformable part model. In Proceedings of the Conference on Computer Vision and Pattern Recognition. 1\u20138."},{"key":"e_1_3_2_37_2","first-page":"1","article-title":"A coarse-to-fine facial landmark detection method based on self-attention mechanism","author":"Lu Jian Xue, Ling Shao, Jiayi Lyu, Pengcheng Gao, and Ke","year":"2020","unstructured":"Jian Xue, Ling Shao, Jiayi Lyu, Pengcheng Gao, and Ke Lu. 2020. A coarse-to-fine facial landmark detection method based on self-attention mechanism. IEEE Trans. Multimedia (2020), 1\u201310.","journal-title":"IEEE Trans. Multimedia"},{"key":"e_1_3_2_38_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.91"},{"key":"e_1_3_2_39_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.690"},{"key":"e_1_3_2_40_2","article-title":"YOLOv3: An incremental improvement","author":"Redmon Joseph","year":"2018","unstructured":"Joseph Redmon and Ali Farhadi. 2018. YOLOv3: An incremental improvement. Retrieved from https:\/\/arXiv:1804.02767.","journal-title":"Retrieved from https:\/\/arXiv:1804.02767"},{"key":"e_1_3_2_41_2","doi-asserted-by":"publisher","DOI":"10.5555\/2969239.2969250"},{"key":"e_1_3_2_42_2","doi-asserted-by":"publisher","DOI":"10.1145\/2393347.2396358"},{"key":"e_1_3_2_43_2","doi-asserted-by":"publisher","DOI":"10.1145\/2461466.2461486"},{"key":"e_1_3_2_44_2","doi-asserted-by":"publisher","DOI":"10.1145\/1991996.1992021"},{"key":"e_1_3_2_45_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2014.81"},{"key":"e_1_3_2_46_2","first-page":"1","volume-title":"Proceedings of the International Conference on Learning Representations","author":"Simonyan Karen","year":"2015","unstructured":"Karen Simonyan and Andrew Zisserman. 2015. Very deep convolutional networks for large-scale image recognition. In Proceedings of the International Conference on Learning Representations. 1\u201314."},{"key":"e_1_3_2_47_2","doi-asserted-by":"publisher","DOI":"10.5555\/3327546.3327604"},{"key":"e_1_3_2_48_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICCVW.2017.41"},{"key":"e_1_3_2_49_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.patcog.2019.107003"},{"key":"e_1_3_2_50_2","first-page":"103","article-title":"Multi-perspective cross-class domain adaptation for open logo detection","author":"Su Hang","year":"2021","unstructured":"Hang Su, Shaogang Gong, and Xiatian Zhu. 2021. Multi-perspective cross-class domain adaptation for open logo detection. Comput. Vision Image Understand. (2021), 103\u2013156.","journal-title":"Comput. Vision Image Understand."},{"key":"e_1_3_2_51_2","doi-asserted-by":"publisher","DOI":"10.1109\/WACV.2017.65"},{"key":"e_1_3_2_52_2","first-page":"111","volume-title":"Proceedings of the British Machine Vision Conference","author":"Su Hang","year":"2018","unstructured":"Hang Su, Xiatian Zhu, and Shaogang Gong. 2018. Open logo detection challenge. In Proceedings of the British Machine Vision Conference. 111\u2013119."},{"key":"e_1_3_2_53_2","article-title":"Sparse r-cnn: End-to-end object detection with learnable proposals","author":"Sun Peize","year":"2020","unstructured":"Peize Sun, Rufeng Zhang, Yi Jiang, Tao Kong, Chenfeng Xu, Wei Zhan, Masayoshi Tomizuka, Lei Li, Zehuan Yuan, Changhu Wang et\u00a0al. 2020. Sparse r-cnn: End-to-end object detection with learnable proposals. Retrieved from https:\/\/arXiv:2011.12450.","journal-title":"Retrieved from https:\/\/arXiv:2011.12450"},{"key":"e_1_3_2_54_2","first-page":"284","volume-title":"Proceedings of the Conference on Computer Vision, Imaging and Computer Graphics Theory and Applications","author":"T\u00fczk\u00f6 Andras","year":"2018","unstructured":"Andras T\u00fczk\u00f6, Christian Herrmann, Daniel Manger, and J\u00fcrgen Beyerer. 2018. Open set logo detection and retrieval. In Proceedings of the Conference on Computer Vision, Imaging and Computer Graphics Theory and Applications. 284\u2013292."},{"key":"e_1_3_2_55_2","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v34i04.6085"},{"key":"e_1_3_2_56_2","doi-asserted-by":"publisher","DOI":"10.1145\/3408299"},{"key":"e_1_3_2_57_2","doi-asserted-by":"publisher","DOI":"10.1007\/s00530-005-0167-6"},{"key":"e_1_3_2_58_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2015.7299023"},{"key":"e_1_3_2_59_2","doi-asserted-by":"publisher","DOI":"10.1145\/3007669.3007728"},{"key":"e_1_3_2_60_2","doi-asserted-by":"publisher","DOI":"10.1145\/2578726.2578748"},{"key":"e_1_3_2_61_2","first-page":"2115","article-title":"Filtering of brand-related microblogs using social-smooth multiview embedding","author":"Zhen Haojie Li, Tat-Seng Chua, Yue Gao, and Yi","year":"2016","unstructured":"Haojie Li, Tat-Seng Chua, Yue Gao, and Yi Zhen. 2016. Filtering of brand-related microblogs using social-smooth multiview embedding. IEEE Trans. Multimedia (2016), 2115\u20132126.","journal-title":"IEEE Trans. Multimedia"},{"key":"e_1_3_2_62_2","doi-asserted-by":"publisher","DOI":"10.1145\/2457450.2457454"},{"key":"e_1_3_2_63_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.00978"},{"key":"e_1_3_2_64_2","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v34i07.6999"},{"key":"e_1_3_2_65_2","volume-title":"Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV\u201920)","author":"Chen Hao","year":"2020","unstructured":"Hao Chen, Tong He, Zhi Tian, and Chunhua Shen. 2020. FCOS: Fully convolutional one-stage object detection. In Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV\u201920)."},{"key":"e_1_3_2_66_2","doi-asserted-by":"publisher","DOI":"10.1109\/TMM.2016.2647386"},{"key":"e_1_3_2_67_2","doi-asserted-by":"publisher","DOI":"10.1145\/3352691"}],"container-title":["ACM Transactions on Multimedia Computing, Communications, and Applications"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3466780","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3466780","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T21:28:09Z","timestamp":1750195689000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3466780"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,1,27]]},"references-count":66,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2022,1,31]]}},"alternative-id":["10.1145\/3466780"],"URL":"https:\/\/doi.org\/10.1145\/3466780","relation":{},"ISSN":["1551-6857","1551-6865"],"issn-type":[{"value":"1551-6857","type":"print"},{"value":"1551-6865","type":"electronic"}],"subject":[],"published":{"date-parts":[[2022,1,27]]},"assertion":[{"value":"2020-09-01","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2021-05-01","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2022-01-27","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}