{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,21]],"date-time":"2026-05-21T05:21:05Z","timestamp":1779340865798,"version":"3.51.4"},"reference-count":73,"publisher":"Springer Science and Business Media LLC","issue":"2","license":[{"start":{"date-parts":[[2012,3,28]],"date-time":"2012-03-28T00:00:00Z","timestamp":1332892800000},"content-version":"tdm","delay-in-days":0,"URL":"http:\/\/www.springer.com\/tdm"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Int J Comput Vis"],"published-print":{"date-parts":[[2012,9]]},"DOI":"10.1007\/s11263-012-0524-9","type":"journal-article","created":{"date-parts":[[2012,3,27]],"date-time":"2012-03-27T12:35:14Z","timestamp":1332851714000},"page":"190-214","source":"Crossref","is-referenced-by-count":176,"title":["2D Articulated Human Pose Estimation and Retrieval in (Almost) Unconstrained Still Images"],"prefix":"10.1007","volume":"99","author":[{"given":"M.","family":"Eichner","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"M.","family":"Marin-Jimenez","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"A.","family":"Zisserman","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"V.","family":"Ferrari","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2012,3,28]]},"reference":[{"key":"524_CR1","volume-title":"CVPR","author":"A. Agarwal","year":"2004","unstructured":"Agarwal, A., & Triggs, B. (2004a). 3d human pose from silhouettes by relevance vector regression. In CVPR."},{"key":"524_CR2","volume-title":"ECCV","author":"A. Agarwal","year":"2004","unstructured":"Agarwal, A., & Triggs, B. (2004b). Tracking articulated motion using a mixture of autoregressive models. In ECCV."},{"key":"524_CR3","volume-title":"CVPR","author":"M. Andriluka","year":"2009","unstructured":"Andriluka, M., Roth, S., & Schiele, B. (2009). Pictorial structures revisited: people detection and articulated pose estimation. In CVPR."},{"key":"524_CR4","volume-title":"CVPR","author":"O. Arandjelovic","year":"2005","unstructured":"Arandjelovic, O., & Zisserman, A. (2005). Automatic face recognition for film character retrieval in feature-length films. In CVPR."},{"key":"524_CR5","volume-title":"DAGM","author":"M. Bergtholdt","year":"2008","unstructured":"Bergtholdt, M., Knappes, J., & Schnorr, C. (2008). Learning of graphical models and efficient inference for object class recognition. In DAGM."},{"key":"524_CR6","volume-title":"Pattern recognition and machine learning","author":"C. Bishop","year":"2006","unstructured":"Bishop, C. (2006). Pattern recognition and machine learning. Berlin: Springer."},{"key":"524_CR7","volume-title":"ICCV","author":"M. Blank","year":"2005","unstructured":"Blank, M., Gorelick, L., Shechtman, E., Irani, M., & Basri, R. (2005). Actions as space-time shapes. In ICCV."},{"key":"524_CR8","volume-title":"PAMI","author":"A. Bobick","year":"2001","unstructured":"Bobick, A., & Davis, J. (2001). The recognition of human movement using temporal templates. In PAMI."},{"key":"524_CR9","volume-title":"BMVC","author":"P. Buehler","year":"2008","unstructured":"Buehler, P., Everinghan, M., Huttenlocher, D., & Zisserman, A. (2008). Long term arm and hand tracking for continuous sign language TV broadcasts. In BMVC."},{"key":"524_CR10","volume-title":"CVPR","author":"T. Cham","year":"1999","unstructured":"Cham, T., & Rehg, J. (1999). A\u00a0multiple hypothesis approach to figure tracking. In CVPR."},{"key":"524_CR11","volume-title":"PAMI","author":"D. Comaniciu","year":"2002","unstructured":"Comaniciu, D., & Meer, P. (2002). Mean shift: a robust approach toward feature space analysis. In PAMI."},{"key":"524_CR12","volume-title":"SIGGRAPH","author":"F. Crow","year":"1984","unstructured":"Crow, F. (1984). Summed-area tables for texture mapping. In SIGGRAPH."},{"key":"524_CR13","volume-title":"CVPR","author":"N. Dalal","year":"2005","unstructured":"Dalal, N., & Triggs, B. (2005). Histogram of oriented gradients for human detection. In CVPR."},{"key":"524_CR14","volume-title":"ICCV VS-PETS","author":"P. Dollar","year":"2005","unstructured":"Dollar, P., Rabaud, V., Cottrell, G., & Belongie, S. (2005). Behavior recognition via sparse spatio-temporal features. In ICCV VS-PETS."},{"key":"524_CR15","volume-title":"BMVC","author":"M. Eichner","year":"2009","unstructured":"Eichner, M., & Ferrari, V. (2009). Better appearance models for pictorial structures. In BMVC."},{"key":"524_CR16","volume-title":"ECCV","author":"M. Eichner","year":"2010","unstructured":"Eichner, M., & Ferrari, V. (2010). We are family: Joint pose estimation of multiple persons. In ECCV."},{"key":"524_CR17","unstructured":"Everingham, M., Van Gool, L., Williams, C. K. I., Winn, J., & Zisserman, A. (2008). The PASCAL visual object classes challenge 2008 (VOC2008) results."},{"key":"524_CR18","volume-title":"CVPR","author":"A. Fathi","year":"2008","unstructured":"Fathi, A., & Mori, G. (2008). Action recognition by learning mid-level motion features. In CVPR."},{"issue":"1","key":"524_CR19","doi-asserted-by":"crossref","first-page":"55","DOI":"10.1023\/B:VISI.0000042934.15159.49","volume":"61","author":"P. Felzenszwalb","year":"2005","unstructured":"Felzenszwalb, P., & Huttenlocher, D. (2005). Pictorial structures for object recognition. International Journal of Computer Vision, 61(1), 55\u201379.","journal-title":"International Journal of Computer Vision"},{"key":"524_CR20","volume-title":"CVPR","author":"P. Felzenszwalb","year":"2008","unstructured":"Felzenszwalb, P., McAllester, D., & Ramanan, D. (2008). A\u00a0discriminatively trained, multiscale, deformable part model. In CVPR."},{"key":"524_CR21","volume-title":"CVPR","author":"V. Ferrari","year":"2001","unstructured":"Ferrari, V., Tuytelaars, T., & Van Gool, L. (2001). Real-time affine region tracking and coplanar grouping. In CVPR."},{"key":"524_CR22","volume-title":"CVPR","author":"V. Ferrari","year":"2008","unstructured":"Ferrari, V., Marin-Jimenez, M., & Zisserman, A. (2008). Progressive search space reduction for human pose estimation. In CVPR."},{"key":"524_CR23","volume-title":"CVPR","author":"V. Ferrari","year":"2009","unstructured":"Ferrari, V., Marin-Jimenez, M., & Zisserman, A. (2009). Pose search: retrieving people using their pose. In CVPR."},{"key":"524_CR24","volume-title":"CVPR","author":"D. Forsyth","year":"1997","unstructured":"Forsyth, D., & Fleck, M. (1997). Body plans. In CVPR."},{"key":"524_CR25","volume-title":"ECCV","author":"D. M. Gavrilla","year":"2000","unstructured":"Gavrilla, D. M. (2000). Pedestrian detection from a moving vehicle. In ECCV."},{"key":"524_CR26","volume-title":"ICCV","author":"P. Guan","year":"2009","unstructured":"Guan, P., Weiss, A., Balan, A., & Black, M. (2009). Estimating human shape and pose from a single image. In ICCV."},{"key":"524_CR27","volume-title":"CVPR","author":"G. Hua","year":"2005","unstructured":"Hua, G., Yang, M. H., & Wu, Y. (2005). Learning to estimate human pose with data driven belief propagation. In CVPR."},{"key":"524_CR28","volume-title":"ICCV workshop on human motion understanding","author":"N. Ikizler","year":"2007","unstructured":"Ikizler, N., & Duygulu, P. (2007). Human action recognition using distribution of oriented rectangular patches. In ICCV workshop on human motion understanding."},{"key":"524_CR29","volume-title":"ICCV","author":"S. Ioffe","year":"1999","unstructured":"Ioffe, S., & Forsyth, D. (1999). Finding people by sampling. In ICCV."},{"key":"524_CR30","volume-title":"ICCV","author":"H. Jiang","year":"2009","unstructured":"Jiang, H. (2009). Human pose estimation using consistent max-covering. In ICCV."},{"key":"524_CR31","volume-title":"CVPR","author":"H. Jiang","year":"2008","unstructured":"Jiang, H., & Martin, D. R. (2008). Global pose estimation using non-tree models. In CVPR."},{"key":"524_CR32","volume-title":"MLVMA","author":"S. Johnson","year":"2009","unstructured":"Johnson, S., & Everingham, M. (2009). Combining discriminative appearance and segmentation cues for articulated human pose estimation. In MLVMA."},{"key":"524_CR33","volume-title":"BMVC","author":"S. Johnson","year":"2010","unstructured":"Johnson, S., & Everingham, M. (2010). Clustered pose and nonlinear appearance models for human pose estimation. In BMVC."},{"key":"524_CR34","volume-title":"CVPR","author":"Y. Ke","year":"2007","unstructured":"Ke, Y., Sukthankar, R., & Hebert, M. (2007). Spatio-temporal shape and flow correlation for action recognition. In CVPR."},{"key":"524_CR35","volume-title":"ICVGIP","author":"M. P. Kumar","year":"2004","unstructured":"Kumar, M. P., Torr, P. H. S., & Zisserman, A. (2004). Learning layered pictorial structures from video. In ICVGIP."},{"key":"524_CR36","volume-title":"ICCV","author":"M. P. Kumar","year":"2009","unstructured":"Kumar, M. P., Torr, P. H. S., & Zisserman, A. (2009). Efficient discriminative learning of parts-based models. In ICCV."},{"key":"524_CR37","volume-title":"CVPR","author":"X. Lan","year":"2004","unstructured":"Lan, X., & Huttenlocher, D. P. (2004). A\u00a0unified spatio-temporal articulated model for tracking. In CVPR."},{"key":"524_CR38","volume-title":"ICCV","author":"X. Lan","year":"2005","unstructured":"Lan, X., & Huttenlocher, D. (2005). Beyond trees: common-factor models for 2D human pose recovery. In ICCV."},{"key":"524_CR39","volume-title":"BMVC","author":"I. Laptev","year":"2006","unstructured":"Laptev, I. (2006). Improvements of object detection using boosted histograms. In BMVC."},{"key":"524_CR40","volume-title":"ICCV","author":"I. Laptev","year":"2007","unstructured":"Laptev, I., Perez, P. (2007). Retrieving actions in movies. In ICCV."},{"key":"524_CR41","volume-title":"CVPR","author":"I. Laptev","year":"2008","unstructured":"Laptev, I., Marsza\u0142ek, M., Schmid, C., & Rozenfeld, B. (2008). Learning realistic human actions from movies. In CVPR."},{"key":"524_CR42","volume-title":"CVPR","author":"M.W. Lee","year":"2004","unstructured":"Lee, M.W., Cohen, I. (2004). Proposal maps driven MCMC for estimating human body pose in static images. In CVPR."},{"key":"524_CR43","volume-title":"CIVR","author":"P. Li","year":"2007","unstructured":"Li, P., Ai, H., Li, Y., & Huang, C. (2007). Video parsing based on head tracking and face recognition. In CIVR."},{"key":"524_CR44","volume-title":"ECCV","author":"K. Mikolajczyk","year":"2004","unstructured":"Mikolajczyk, K., Schmid, C., & Zisserman, A. (2004). Human detection based on a probabilistic assembly of robust part detectors. In ECCV."},{"key":"524_CR45","volume-title":"CVPR","author":"G. Mori","year":"2002","unstructured":"Mori, G., & Malik, J. (2002). Estimating human body configurations using shape context matching. In CVPR."},{"key":"524_CR46","volume-title":"CVPR","author":"J. Niebles","year":"2007","unstructured":"Niebles, J., & Fei-Fei, L. (2007). A\u00a0hierarchical model of shape and appearance for human action classification. In CVPR."},{"key":"524_CR47","volume-title":"Numerical optimization","author":"J. Nocedal","year":"2006","unstructured":"Nocedal, J., & Wright, S. (2006). Numerical optimization. Berlin: Springer."},{"key":"524_CR48","volume-title":"ECCV","author":"M. Ozuysal","year":"2006","unstructured":"Ozuysal, M., Lepetit, V., Fleuret, F., & Fua, P. (2006). Feature harvesting for tracking-by-detection. In ECCV."},{"key":"524_CR49","volume-title":"NIPS","author":"D. Ramanan","year":"2006","unstructured":"Ramanan, D. (2006). Learning to parse images of articulated bodies. In NIPS."},{"key":"524_CR50","volume-title":"CVPR","author":"D. Ramanan","year":"2005","unstructured":"Ramanan, D., Forsyth, D. A., & Zisserman, A. (2005). Strike a pose: tracking people by finding stylized poses. In CVPR."},{"key":"524_CR51","volume-title":"CVPR","author":"X. Ren","year":"2005","unstructured":"Ren, X., Berg, A., & Malik, J. (2005). Recovering human body configurations using pairwise constraints between parts. In CVPR."},{"key":"524_CR52","volume-title":"ECCV","author":"R. Ronfard","year":"2002","unstructured":"Ronfard, R., Schmid, C., & Triggs, B. (2002). Learning to parse pictures of people. In ECCV."},{"key":"524_CR53","volume-title":"SIGGRAPH","author":"C. Rother","year":"2004","unstructured":"Rother, C., Kolmogorov, V., & Blake, A. (2004). Grabcut: interactive foreground extraction using iterated graph cuts. In SIGGRAPH."},{"key":"524_CR54","volume-title":"CVPR","author":"B. Sapp","year":"2010","unstructured":"Sapp, B., Jordan, C., & Taskar, B. (2010a). Adaptive pose priors for pictorial structures. In CVPR."},{"key":"524_CR55","volume-title":"ECCV","author":"B. Sapp","year":"2010","unstructured":"Sapp, B., Toshev, A., & Taskar, B. (2010b). Cascaded models for articulated pose estimation. In ECCV."},{"key":"524_CR56","volume-title":"CVPR","author":"E. Shechtman","year":"2007","unstructured":"Shechtman, E., & Irani, M. (2007). Matching local self-similarities across images and videos. In CVPR."},{"key":"524_CR57","volume-title":"CVPR","author":"L. Sigal","year":"2006","unstructured":"Sigal, L., & Black, M. (2006). Measure locally, reason globally: occlusion-sensitive articulated pose estimation. In CVPR."},{"key":"524_CR58","volume-title":"NIPS","author":"L. Sigal","year":"2003","unstructured":"Sigal, L., Isard, M., Sigelman, B. H., & Black, M. J. (2003). Attractive people: assembling loose-limbed models using non-parametric belief propagation. In NIPS."},{"key":"524_CR59","volume-title":"ECCV","author":"V. K. Singh","year":"2010","unstructured":"Singh, V. K., Nevatia, R., & Huang, C. (2010). Efficient inference with multiple heterogeneous part detectors for human pose estimation. In ECCV."},{"key":"524_CR60","volume-title":"ICCV","author":"J. Sivic","year":"2003","unstructured":"Sivic, J., & Zisserman, A. (2003). Video Google: a text retrieval approach to object matching in videos. In ICCV."},{"key":"524_CR61","volume-title":"CIVR","author":"J. Sivic","year":"2005","unstructured":"Sivic, J., Everingham, M., & Zisserman, A. (2005). Person spotting: video shot retrieval for face sets. In CIVR."},{"key":"524_CR62","volume-title":"CVPR","author":"T. P. Tian","year":"2010","unstructured":"Tian, T. P., & Sclaroff, S. (2010a). Fast globally optimal 2D human detection with loopy graph models. In CVPR."},{"key":"524_CR63","volume-title":"ECCV","author":"T. P. Tian","year":"2010","unstructured":"Tian, T. P., & Sclaroff, S. (2010b). Fast multi-aspect 2D human detection. In ECCV."},{"key":"524_CR64","volume-title":"ECCV","author":"D. Tran","year":"2010","unstructured":"Tran, D., & Forsyth, D. (2010). Improved human parsing with a full relational model. In ECCV."},{"key":"524_CR65","volume-title":"CVPR","author":"P. Viola","year":"2001","unstructured":"Viola, P., & Jones, M. (2001). Rapid object detection using a boosted cascade of simple features. In CVPR."},{"key":"524_CR66","volume-title":"ECCV","author":"Y. Wang","year":"2008","unstructured":"Wang, Y., & Mori, G. (2008). Multiple tree models for occlusion and spatial constraints in human pose estimation. In ECCV."},{"key":"524_CR67","unstructured":"website (2008). VGG upper body detector. http:\/\/www.robots.ox.ac.uk\/~vgg\/software\/UpperBody\/ ."},{"key":"524_CR68","unstructured":"website (2009a). Buffy stickmen dataset. http:\/\/www.robots.ox.ac.uk\/~vgg\/data\/stickmen\/ ."},{"key":"524_CR69","unstructured":"website (2009b). ETHZ PASCAL stickmen dataset. http:\/\/www.vision.ee.ethz.ch\/~calvin\/ethz_pascal_stickmen\/ ."},{"key":"524_CR70","unstructured":"website (2009c). HPE software. http:\/\/www.vision.ee.ethz.ch\/~calvin\/articulated_human_pose_estimation_code\/ ."},{"key":"524_CR71","unstructured":"website (2009d). VGG pose estimation and search. http:\/\/www.robots.ox.ac.uk\/~vgg\/research\/pose_estimation\/ ."},{"key":"524_CR72","unstructured":"website (2010a). CALVIN upper body detector. http:\/\/www.vision.ee.ethz.ch\/~calvin\/calvin_upperbody_detector\/ ."},{"key":"524_CR73","unstructured":"website (2010b). HPE online demo. http:\/\/www.vision.ee.ethz.ch\/~hpedemo\/ ."}],"container-title":["International Journal of Computer Vision"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/link.springer.com\/content\/pdf\/10.1007\/s11263-012-0524-9.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/link.springer.com\/article\/10.1007\/s11263-012-0524-9\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/link.springer.com\/content\/pdf\/10.1007\/s11263-012-0524-9","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2019,6,1]],"date-time":"2019-06-01T08:16:48Z","timestamp":1559377008000},"score":1,"resource":{"primary":{"URL":"http:\/\/link.springer.com\/10.1007\/s11263-012-0524-9"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2012,3,28]]},"references-count":73,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2012,9]]}},"alternative-id":["524"],"URL":"https:\/\/doi.org\/10.1007\/s11263-012-0524-9","relation":{},"ISSN":["0920-5691","1573-1405"],"issn-type":[{"value":"0920-5691","type":"print"},{"value":"1573-1405","type":"electronic"}],"subject":[],"published":{"date-parts":[[2012,3,28]]}}}