{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,17]],"date-time":"2026-07-17T06:02:25Z","timestamp":1784268145195,"version":"3.55.0"},"reference-count":32,"publisher":"Association for Computing Machinery (ACM)","issue":"4","license":[{"start":{"date-parts":[[2012,7,1]],"date-time":"2012-07-01T00:00:00Z","timestamp":1341100800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/100000001","name":"National Science Foundation","doi-asserted-by":"publisher","award":["1149853"],"award-info":[{"award-number":["1149853"]}],"id":[{"id":"10.13039\/100000001","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Graph."],"published-print":{"date-parts":[[2012,8,5]]},"abstract":"<jats:p>Humans have used sketching to depict our visual world since prehistoric times. Even today, sketching is possibly the only rendering technique readily available to all humans. This paper is the first large scale exploration of human sketches. We analyze the distribution of non-expert sketches of everyday objects such as 'teapot' or 'car'. We ask humans to sketch objects of a given category and gather 20,000 unique sketches evenly distributed over 250 object categories. With this dataset we perform a perceptual study and find that humans can correctly identify the object category of a sketch 73% of the time. We compare human performance against computational recognition methods. We develop a bag-of-features sketch representation and use multi-class support vector machines, trained on our sketch dataset, to classify sketches. The resulting recognition method is able to identify unknown sketches with 56% accuracy (chance is 0.4%). Based on the computational model, we demonstrate an interactive sketch recognition system. We release the complete crowd-sourced dataset of sketches to the community.<\/jats:p>","DOI":"10.1145\/2185520.2185540","type":"journal-article","created":{"date-parts":[[2012,8,6]],"date-time":"2012-08-06T18:11:37Z","timestamp":1344276697000},"page":"1-10","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":527,"title":["How do humans sketch objects?"],"prefix":"10.1145","volume":"31","author":[{"given":"Mathias","family":"Eitz","sequence":"first","affiliation":[{"name":"TU Berlin"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"James","family":"Hays","sequence":"additional","affiliation":[{"name":"Brown University"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Marc","family":"Alexa","sequence":"additional","affiliation":[{"name":"TU Berlin"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2012,7]]},"reference":[{"key":"e_1_2_2_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/TSMCA.2004.838464"},{"key":"e_1_2_2_2_1","doi-asserted-by":"publisher","DOI":"10.1145\/1618452.1618470"},{"key":"e_1_2_2_3_1","doi-asserted-by":"publisher","DOI":"10.1145\/1348246.1348248"},{"key":"e_1_2_2_4_1","doi-asserted-by":"publisher","DOI":"10.1145\/1753326.1753459"},{"key":"e_1_2_2_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/TVCG.2010.266"},{"key":"e_1_2_2_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/MCG.2011.67"},{"key":"e_1_2_2_7_1","doi-asserted-by":"publisher","DOI":"10.1145\/2185520.2185527"},{"key":"e_1_2_2_8_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-009-0275-4"},{"key":"e_1_2_2_9_1","doi-asserted-by":"publisher","DOI":"10.1145\/2070781.2024167"},{"key":"e_1_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1145\/258734.258849"},{"key":"e_1_2_2_11_1","volume-title":"Conf. Computer Vision, 456--463","author":"Georgescu B.","unstructured":"Georgescu , B. , Shimshoni , I. , and Meer , P . 2003. Mean shift based clustering in high dimensions: a texture classification example. in IEEE Int'l . Conf. Computer Vision, 456--463 . Georgescu, B., Shimshoni, I., and Meer, P. 2003. Mean shift based clustering in high dimensions: a texture classification example. in IEEE Int'l. Conf. Computer Vision, 456--463."},{"key":"e_1_2_2_12_1","unstructured":"Griffin G. Holub A. and Perona P. 2007. Caltech-256 object category dataset. Tech. rep. California institute of Technology. Griffin G. Holub A. and Perona P. 2007. Caltech-256 object category dataset. Tech. rep. California institute of Technology."},{"key":"e_1_2_2_13_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.cag.2005.05.005"},{"key":"e_1_2_2_14_1","doi-asserted-by":"publisher","DOI":"10.1145\/563274.563294"},{"key":"e_1_2_2_15_1","doi-asserted-by":"publisher","DOI":"10.1145\/1015706.1015741"},{"key":"e_1_2_2_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2006.68"},{"key":"e_1_2_2_17_1","doi-asserted-by":"publisher","DOI":"10.1145\/2010324.1964922"},{"key":"e_1_2_2_18_1","doi-asserted-by":"publisher","DOI":"10.1023\/B:VISI.0000029664.99615.94"},{"key":"e_1_2_2_19_1","doi-asserted-by":"publisher","DOI":"10.1145\/1943403.1943444"},{"key":"e_1_2_2_20_1","doi-asserted-by":"publisher","DOI":"10.1145\/1378773.1378775"},{"key":"e_1_2_2_21_1","volume-title":"IEEE Conf. Computer Vision and Pattern Recognition, 1--8.","author":"Philbin J.","unstructured":"Philbin , J. , Chum , O. , Isard , M. , Sivic , J. , and Zisserman , A . 2008. Lost in quantization: improving particular object retrieval in large scale image databases . In IEEE Conf. Computer Vision and Pattern Recognition, 1--8. Philbin, J., Chum, O., Isard, M., Sivic, J., and Zisserman, A. 2008. Lost in quantization: improving particular object retrieval in large scale image databases. In IEEE Conf. Computer Vision and Pattern Recognition, 1--8."},{"key":"e_1_2_2_22_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-007-0090-8"},{"key":"e_1_2_2_23_1","unstructured":"Samet H. 2006. Foundations of multidimensional and metric data structures. Morgan Kaufmann. Samet H. 2006. Foundations of multidimensional and metric data structures . Morgan Kaufmann."},{"key":"e_1_2_2_24_1","doi-asserted-by":"crossref","unstructured":"Sch\u00f6lkopf B. and Smola A. 2002. Learning with kernels. MIT Press. Sch\u00f6lkopf B. and Smola A. 2002. Learning with kernels . MIT Press.","DOI":"10.7551\/mitpress\/4175.001.0001"},{"key":"e_1_2_2_25_1","doi-asserted-by":"publisher","DOI":"10.1145\/971478.971487"},{"key":"e_1_2_2_26_1","doi-asserted-by":"publisher","DOI":"10.5555\/998687.1007045"},{"key":"e_1_2_2_27_1","doi-asserted-by":"publisher","DOI":"10.1145\/2070781.2024188"},{"key":"e_1_2_2_28_1","volume-title":"Conf. Computer Vision, 1470--1477","author":"Sivic J.","unstructured":"Sivic , J. , and Zisserman , A . 2003. Video Google: a textretrieval approach to object matching in videos. In IEEE Int'l . Conf. Computer Vision, 1470--1477 . Sivic, J., and Zisserman, A. 2003. Video Google: a textretrieval approach to object matching in videos. In IEEE Int'l. Conf. Computer Vision, 1470--1477."},{"key":"e_1_2_2_29_1","doi-asserted-by":"publisher","DOI":"10.1145\/1461551.1461591"},{"key":"e_1_2_2_30_1","first-page":"2579","article-title":"Visualizing data using t-SNE","volume":"9","author":"van der Maaten L.","year":"2008","unstructured":"van der Maaten , L. , and Hinton , G. 2008 . Visualizing data using t-SNE . Journal of Machine Learning Research 9 , 2579 -- 2605 . van der Maaten, L., and Hinton, G. 2008. Visualizing data using t-SNE. Journal of Machine Learning Research 9, 2579--2605.","journal-title":"Journal of Machine Learning Research"},{"key":"e_1_2_2_31_1","doi-asserted-by":"publisher","DOI":"10.1073\/pnas.1015666108"},{"key":"e_1_2_2_32_1","volume-title":"Proc. IEEE Conf. Computer Vision and Pattern Recognition, 3485--3492","author":"Xiao J.","unstructured":"Xiao , J. , Hays , J. , Ehinger , K. A. , Oliva , A. , and Torralba , A . 2010. SUN database: large-scale scene recognition from abbey to zoo . In Proc. IEEE Conf. Computer Vision and Pattern Recognition, 3485--3492 . Xiao, J., Hays, J., Ehinger, K. A., Oliva, A., and Torralba, A. 2010. SUN database: large-scale scene recognition from abbey to zoo. In Proc. IEEE Conf. Computer Vision and Pattern Recognition, 3485--3492."}],"container-title":["ACM Transactions on Graphics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2185520.2185540","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/2185520.2185540","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T10:06:47Z","timestamp":1750241207000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2185520.2185540"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2012,7]]},"references-count":32,"journal-issue":{"issue":"4","published-print":{"date-parts":[[2012,8,5]]}},"alternative-id":["10.1145\/2185520.2185540"],"URL":"https:\/\/doi.org\/10.1145\/2185520.2185540","relation":{},"ISSN":["0730-0301","1557-7368"],"issn-type":[{"value":"0730-0301","type":"print"},{"value":"1557-7368","type":"electronic"}],"subject":[],"published":{"date-parts":[[2012,7]]},"assertion":[{"value":"2012-07-01","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}