{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,9]],"date-time":"2026-04-09T10:28:09Z","timestamp":1775730489836,"version":"3.50.1"},"reference-count":32,"publisher":"Association for Computing Machinery (ACM)","issue":"9","license":[{"start":{"date-parts":[[2013,9,1]],"date-time":"2013-09-01T00:00:00Z","timestamp":1377993600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/100000145","name":"Division of Information and Intelligent Systems","doi-asserted-by":"publisher","award":["IIS 0746569, IIS 0811340, IIS 0812428"],"award-info":[{"award-number":["IIS 0746569, IIS 0811340, IIS 0812428"]}],"id":[{"id":"10.13039\/100000145","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["Commun. ACM"],"published-print":{"date-parts":[[2013,9]]},"abstract":"<jats:p>We describe a state-of-the-art system for finding objects in cluttered images. Our system is based on deformable models that represent objects using local part templates and geometric constraints on the locations of parts. We reduce object detection to classification with latent variables. The latent variables introduce invariances that make it possible to detect objects with highly variable appearance. We use a generalization of support vector machines to incorporate latent information during training. This has led to a general framework for discriminative training of classifiers with latent variables. Discriminative training benefits from large training datasets. In practice we use an iterative algorithm that alternates between estimating latent values for positive examples and solving a large convex optimization problem. Practical optimization of this large convex problem can be done using active set techniques for adaptive subsampling of the training data.<\/jats:p>","DOI":"10.1145\/2494532","type":"journal-article","created":{"date-parts":[[2020,4,4]],"date-time":"2020-04-04T11:15:14Z","timestamp":1585998914000},"page":"97-105","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":81,"title":["Visual object detection with deformable part models"],"prefix":"10.1145","volume":"56","author":[{"given":"Pedro","family":"Felzenszwalb","sequence":"first","affiliation":[{"name":"Brown University"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Ross","family":"Girshick","sequence":"additional","affiliation":[{"name":"EECS, UC Berkeley"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"David","family":"McAllester","sequence":"additional","affiliation":[{"name":"Toyota Technological Institute at Chicago"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Deva","family":"Ramanan","sequence":"additional","affiliation":[{"name":"UC Irvine"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2013,9]]},"reference":[{"key":"e_1_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-006-0033-9"},{"key":"e_1_2_1_2_1","volume-title":"Advances in Neural Information Processing Systems","volume":"15","author":"Andrews S.","year":"2003"},{"key":"e_1_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.5555\/645312.648915"},{"key":"e_1_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1109\/34.927467"},{"key":"e_1_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1006\/cviu.2000.0842"},{"key":"e_1_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2005.329"},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2005.177"},{"key":"e_1_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-011-0439-x"},{"key":"e_1_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-009-0275-4"},{"key":"e_1_2_1_10_1","doi-asserted-by":"crossref","unstructured":"Felzenszwalb P. Girshick R. McAllester D. Cascade object detection with deformable part models. In IEEE Computer Vision and Pattern Recognition (2010).  Felzenszwalb P. Girshick R. McAllester D. Cascade object detection with deformable part models. In IEEE Computer Vision and Pattern Recognition (2010).","DOI":"10.1109\/CVPR.2010.5539906"},{"key":"e_1_2_1_11_1","unstructured":"Felzenszwalb P. Girshick R. McAllester D. Ramanan D. Discriminatively trained deformable part models. http:\/\/people.cs.uchicago.edu\/~pff\/latent\/.  Felzenszwalb P. Girshick R. McAllester D. Ramanan D. Discriminatively trained deformable part models. http:\/\/people.cs.uchicago.edu\/~pff\/latent\/."},{"key":"e_1_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2009.167"},{"key":"e_1_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1023\/B:VISI.0000042934.15159.49"},{"key":"e_1_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2008.4587597"},{"key":"e_1_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2003.1211479"},{"key":"e_1_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/T-C.1973.223602"},{"key":"e_1_2_1_19_1","volume-title":"Advances in Neural Information Processing Systems","volume":"24","author":"Girshick R.","year":"2011"},{"key":"e_1_2_1_20_1","volume-title":"Springer-Verlag","author":"Grenander U.","year":"1991"},{"key":"e_1_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1109\/34.232073"},{"key":"e_1_2_1_22_1","volume-title":"IEEE International Conference on Computer Vision","author":"Lamdan Y.","year":"1988"},{"key":"e_1_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1109\/5.726791"},{"key":"e_1_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1016\/0004-3702(87)90070-1"},{"key":"e_1_2_1_25_1","first-page":"1140","volume":"200","author":"Marr D.","year":"1978","journal-title":"Proc. Roy. Soc. Lond. B Biol. Sci."},{"key":"e_1_2_1_26_1","unstructured":"Mundy J. Zisserman A. etal Geometric Invariance in Computer Vision volume 92 MIT press Cambridge MA 1992.   Mundy J. Zisserman A. et al. Geometric Invariance in Computer Vision volume 92 MIT press Cambridge MA 1992."},{"key":"e_1_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1007\/BF01421486"},{"key":"e_1_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2000.855895"},{"key":"e_1_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1109\/34.655648"},{"key":"e_1_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1023\/B:VISI.0000013087.49260.fb"},{"key":"e_1_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2000.854754"},{"key":"e_1_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2011.5995741"},{"key":"e_1_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.1007\/BF00127169"},{"key":"e_1_2_1_34_1","volume-title":"IEEE Conference on Computer Vision and Pattern Recognition","author":"Zhu X.","year":"2012"}],"container-title":["Communications of the ACM"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2494532","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/2494532","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T07:28:55Z","timestamp":1750231735000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2494532"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2013,9]]},"references-count":32,"aliases":["10.1145\/2500468.2494532","10.1145\/2500468.2494532"],"journal-issue":{"issue":"9","published-print":{"date-parts":[[2013,9]]}},"alternative-id":["10.1145\/2494532"],"URL":"https:\/\/doi.org\/10.1145\/2494532","relation":{},"ISSN":["0001-0782","1557-7317"],"issn-type":[{"value":"0001-0782","type":"print"},{"value":"1557-7317","type":"electronic"}],"subject":[],"published":{"date-parts":[[2013,9]]},"assertion":[{"value":"2013-09-01","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}