{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,10,21]],"date-time":"2025-10-21T15:00:31Z","timestamp":1761058831101,"version":"3.35.0"},"reference-count":62,"publisher":"Springer Science and Business Media LLC","issue":"1","content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Int J Comput Vis"],"published-print":{"date-parts":[[2008,10]]},"DOI":"10.1007\/s11263-008-0139-3","type":"journal-article","created":{"date-parts":[[2008,5,12]],"date-time":"2008-05-12T20:13:08Z","timestamp":1210623188000},"page":"16-44","source":"Crossref","is-referenced-by-count":72,"title":["Learning an Alphabet of Shape and Appearance for Multi-Class Object Detection"],"prefix":"10.1007","volume":"80","author":[{"given":"Andreas","family":"Opelt","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Axel","family":"Pinz","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Andrew","family":"Zisserman","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2008,5,13]]},"reference":[{"issue":"11","key":"139_CR1","doi-asserted-by":"crossref","first-page":"1475","DOI":"10.1109\/TPAMI.2004.108","volume":"26","author":"S. Agarwal","year":"2004","unstructured":"Agarwal, S., Awan, A., & Roth, D. (2004). Learning to detect objects in images via a sparse, part-based representation. IEEE Transactions on Pattern Analysis and Machine Intelligence, 26(11), 1475\u20131490.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"139_CR2","doi-asserted-by":"crossref","first-page":"1606","DOI":"10.1109\/TPAMI.2004.111","volume":"28","author":"Y. Amit","year":"2004","unstructured":"Amit, Y., German, D., & Fan, X. (2004). A coarse-to-fine strategy for multi-class shape detection. IEEE Transactions on Pattern Analysis and Machine Intelligence, 28, 1606\u20131621.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"139_CR3","doi-asserted-by":"crossref","unstructured":"Amores, J., Sebe, N., & Radeva, P. (2005). Fast spatial pattern discovery integrating boosting with constellations of contextual descriptors. In Proceedings of the CVPR (Vol. 2, pp. 769\u2013774), CA, USA, June 2005.","DOI":"10.1109\/CVPR.2005.156"},{"key":"139_CR4","doi-asserted-by":"crossref","unstructured":"Bar-Hillel, A., Hertz, T., & Weinshall, D. (2005). Object class recognition by boosting a part-based model. In Proceedings of the CVPR (Vol. 1, pp. 702\u2013709), June 2005.","DOI":"10.1109\/CVPR.2005.250"},{"key":"139_CR5","unstructured":"Bart, E., & Ullman, S. (2005). Cross-generalization:learning novel classes from a single example by feature replacement. In Proceedings of the CVPR (Vol. 1, pp. 672\u2013679)."},{"key":"139_CR6","doi-asserted-by":"crossref","unstructured":"Bernstein, E. J., & Amit, Y. (2005). Part-based statistical models for object classification and detection. In Proceedings of the CVPR (Vol. 2, pp. 734\u2013740).","DOI":"10.1109\/CVPR.2005.270"},{"issue":"6","key":"139_CR7","doi-asserted-by":"crossref","first-page":"849","DOI":"10.1109\/34.9107","volume":"10","author":"G. Borgefors","year":"1988","unstructured":"Borgefors, G. (1988). Hierarchical chamfer matching: a parametric edge matching algorithm. IEEE Transactions on Pattern Analysis and Machine Intelligence, 10(6), 849\u2013865.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"issue":"5","key":"139_CR8","doi-asserted-by":"crossref","first-page":"529","DOI":"10.1109\/34.391389","volume":"17","author":"H. Breu","year":"1995","unstructured":"Breu, H., Gil, J., Kirkpatrick, D., & Werman, M. (1995). Linear time Euclidean distance transform algorithms. IEEE Transactions on Pattern Analysis and Machine Intelligence, 17(5), 529\u2013533.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"139_CR9","doi-asserted-by":"crossref","unstructured":"Caputo, B., Wallraven, C., & Nilsback, M. E. (2004). Object categorization via local kernels. In Proceedings of the ICPR (Vol. 2, pp.\u00a0132\u2013135).","DOI":"10.1109\/ICPR.2004.1334079"},{"issue":"5","key":"139_CR10","doi-asserted-by":"crossref","first-page":"603","DOI":"10.1109\/34.1000236","volume":"24","author":"D. Comaniciu","year":"2002","unstructured":"Comaniciu, D., & Meer, P. (2002). Mean shift: a robust approach towards feature space analysis. IEEE Transactions on Pattern Analysis and Machine Intelligence, 24(5), 603\u2013619.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"139_CR11","doi-asserted-by":"crossref","unstructured":"Crandall, D., Felzenszwalb, P., & Huttenlocher, D. (2005). Spatial priors for part-based recognition using statistical models. In Proceedings of the CVPR (pp. 10\u201317).","DOI":"10.1109\/CVPR.2005.329"},{"key":"139_CR12","unstructured":"Csurka, G., Bray, C., Dance, C., & Fan, L. (2004). Visual categorization with bags of keypoints. In ECCV\u201904: workshop on statistical learning in computer vision (pp. 59\u201374)."},{"key":"139_CR13","doi-asserted-by":"crossref","unstructured":"Dalal, N., & Triggs, B. (2005). Histograms of oriented gradients for human detection. In Proceedings of the CVPR (Vol.\u00a01, pp. 886\u2013893).","DOI":"10.1109\/CVPR.2005.177"},{"key":"139_CR14","doi-asserted-by":"crossref","unstructured":"Deselaers, T., Keysers, D., & Ney, H. (2005). Discriminative training for object recognition using images patches. In Proceedings of the CVPR (Vol. 2, pp. 157\u2013162).","DOI":"10.1109\/CVPR.2005.134"},{"key":"139_CR15","unstructured":"Epstein, B., & Ullman, S. (2005). Feature hierarchies for object classification. In Proceedings of the ICCV (Vol. 1, pp. 220\u2013227)."},{"key":"139_CR16","series-title":"Lecture notes in artificial intelligence","volume-title":"Selected proceedings of the first PASCAL challenges workshop","author":"M. Everingham","year":"2005","unstructured":"Everingham, M., Zisserman, A., Williams, C., Van Gool, L., Allan, M., Bishop, C., Chapelle, O., Dalal, N., Deselaers, T., Dorko, G., Duffner, S., Eichhorn, J., Farquhar, J., Fritz, M., Garcia, C., Griffiths, T., Jurie, F., Keysers, D., Koskela, M., Laaksonen, J., Larlus, D., Leibe, B., Meng, H., Ney, H., Schiele, B., Schmid, C., Seemann, E., Shawe-Taylor, J., Storkey, A., Szedmak, S., Triggs, B., Ulusoy, I., Viitaniemi, V., & Zhang, J. (2005). The 2005 pascal visual object classes challenge. In Lecture notes in artificial intelligence. Selected proceedings of the first PASCAL challenges workshop. Berlin: Springer."},{"key":"139_CR17","unstructured":"Fan, X. (2005). Efficient multiclass object detection by a hierarchy of classifiers. In Proceedings of the CVPR (Vol. 1, pp. 716\u2013723)."},{"key":"139_CR18","doi-asserted-by":"crossref","unstructured":"Fei-Fei, L., Fergus, R., & Perona, P. (2004). Learning generative visual models from few training examples: an incremental Bayesian approach tested on 101 object categories. In Proceedings of the CVPR workshop on generative-model based vision.","DOI":"10.1109\/CVPR.2004.383"},{"issue":"1","key":"139_CR19","doi-asserted-by":"crossref","first-page":"55","DOI":"10.1023\/B:VISI.0000042934.15159.49","volume":"61","author":"P. Felzenszwalb","year":"2004","unstructured":"Felzenszwalb, P., & Huttenlocher, D. (2004). Pictorial structures for object recognition. International Journal of Computer Vision, 61(1), 55\u201379.","journal-title":"International Journal of Computer Vision"},{"key":"139_CR20","doi-asserted-by":"crossref","unstructured":"Fergus, R., Perona, P., & Zisserman, A. (2003). Object class recognition by unsupervised scale-invariant learning. In Proceedings of the CVPR (pp. 264\u2013271).","DOI":"10.1109\/CVPR.2003.1211479"},{"key":"139_CR21","doi-asserted-by":"crossref","unstructured":"Fergus, R., Perona, P., & Zisserman, A. (2004). A visual category filter for Google images. In Proceedings of the ECCV (pp. 242\u2013256).","DOI":"10.1007\/978-3-540-24670-1_19"},{"key":"139_CR22","doi-asserted-by":"crossref","unstructured":"Fergus, R., Perona, P., & Zisserman, A. (2005). A sparse object category model for efficient learning and exhaustive recognition. In Proceedings of the CVPR (Vol. 1, pp. 380\u2013387).","DOI":"10.1109\/CVPR.2005.47"},{"issue":"3","key":"139_CR23","doi-asserted-by":"crossref","first-page":"273","DOI":"10.1007\/s11263-006-8707-x","volume":"71","author":"R. Fergus","year":"2007","unstructured":"Fergus, R., Perona, P., & Zisserman, A. (2007). Weakly supervised scale-invariant learning of models for visual recognition. International Journal of Computer Vision, 71(3), 273\u2013303.","journal-title":"International Journal of Computer Vision"},{"key":"139_CR24","doi-asserted-by":"crossref","unstructured":"Ferrari, V., Tuytelaars, T., & Van Gool, L. (2004). Simultaneous object recognition and segmentation by image exploration. In Proceedings of the ECCV (pp. 40\u201354).","DOI":"10.1007\/978-3-540-24670-1_4"},{"key":"139_CR25","doi-asserted-by":"crossref","unstructured":"Ferrari, V., Tuytelaars, T., & Van Gool, L. (2006). Object detection by contour segment networks. In Proceedings of the ECCV (Vol. 3, pp. 14\u201328).","DOI":"10.1007\/11744078_2"},{"issue":"1","key":"139_CR26","doi-asserted-by":"crossref","first-page":"119","DOI":"10.1006\/jcss.1997.1504","volume":"55","author":"Y. Freund","year":"1997","unstructured":"Freund, Y., & Schapire, R. (1997). A decision theoretic generalisation of online learning. Computer and System Sciences, 55(1), 119\u2013139.","journal-title":"Computer and System Sciences"},{"key":"139_CR27","unstructured":"Friedman, J., Hastie, T., & Tibshirani, R. (1998). Additive logistic regression: a statistical view of boosting (Technical report). Stanford University, Department of Statistics, California 94305."},{"key":"139_CR28","doi-asserted-by":"crossref","unstructured":"Gavrila, D. M., & Philomin, V. (1999). Real-time object detection for smart vehicles. In Proceedings of the ICCV (pp. 87\u201393).","DOI":"10.1109\/ICCV.1999.791202"},{"key":"139_CR29","doi-asserted-by":"crossref","unstructured":"Jurie, F., & Schmid, C. (2004). Scale-invariant shape features for recognition of object categories. In Proceedings of conference on vision and pattern recognition (pp. 90\u201396).","DOI":"10.1109\/CVPR.2004.1315149"},{"key":"139_CR30","unstructured":"Kumar, M. P., Torr, P. H. S., & Zisserman, A. (2004). Extending pictural structures for object recognition. In Proceedings of the BMVC."},{"key":"139_CR31","unstructured":"Leibe, B., Leonardis, A., & Schiele, B. (2004). Combined object categorization and segmentation with an implicit shape model. In ECCV\u201904: workshop on statistical learning in computer vision (pp. 17\u201332), May 2004."},{"key":"139_CR32","doi-asserted-by":"crossref","unstructured":"Leibe, B., & Schiele, B. (2004). Scale-invariant object categorization using a scale-adaptive means-shift search. In DAGM\u201904 (pp. 145\u2013153), August 2004.","DOI":"10.1007\/978-3-540-28649-3_18"},{"key":"139_CR33","doi-asserted-by":"crossref","unstructured":"Lowe, D. G. (1999). Object recognition from local scale-invariant features. In Proceedings of the ICCV (pp. 1150\u20131157).","DOI":"10.1109\/ICCV.1999.790410"},{"issue":"8","key":"139_CR34","doi-asserted-by":"crossref","first-page":"581","DOI":"10.1016\/S0262-8856(02)00047-1","volume":"20","author":"D. Magee","year":"2002","unstructured":"Magee, D., & Boyle, R. (2002). Detection of lameness using re-sampling condensation and multi-steam cyclic hidden Markov models. Image and Vision Computing, 20(8), 581\u2013594.","journal-title":"Image and Vision Computing"},{"key":"139_CR35","unstructured":"Marszalek, M., & Schmid, C. (2006). Spatial weighting for bag-of-features. In Proceedings of the CVPR."},{"key":"139_CR36","doi-asserted-by":"crossref","unstructured":"Mikolajczyk, K., Leibe, B., & Schiele, B. (2006). Multiple object class detection with a generative model. In Proceedings of the CVPR.","DOI":"10.1109\/CVPR.2006.202"},{"key":"139_CR37","doi-asserted-by":"crossref","unstructured":"Mutch, J., & Lowe, D. (2006). Multiclass object recognition with sparse, localized features. In Proceedings of the CVPR.","DOI":"10.1109\/CVPR.2006.200"},{"key":"139_CR38","doi-asserted-by":"crossref","unstructured":"Nist\u00e9r, D., & Stew\u00e9nius, H. (2006). Scalable recognition with a vocabulary tree. In Proceedings of the CVPR.","DOI":"10.1109\/CVPR.2006.264"},{"key":"139_CR39","doi-asserted-by":"crossref","unstructured":"Ommer, B., & Buhmann, J. M. (2006). Learning compositional categorization models. In Proceedings of the ECCV (Vol. 3, pp. 316\u2013329).","DOI":"10.1007\/11744078_25"},{"key":"139_CR40","doi-asserted-by":"crossref","unstructured":"Opelt, A., Fussenegger, M., Pinz, A., & Auer, P. (2004). Weak hypotheses and boosting for generic object detection and recognition. In Proceedings of the ECCV (pp. 71\u201384).","DOI":"10.1007\/978-3-540-24671-8_6"},{"key":"139_CR41","doi-asserted-by":"crossref","unstructured":"Opelt, A., Fussenegger, M., Pinz, A., & Auer, P. (2006a). Generic object recognition with boosting. Pattern Analysis and Machine Intelligence, 28(3).","DOI":"10.1109\/TPAMI.2006.54"},{"key":"139_CR42","doi-asserted-by":"crossref","unstructured":"Opelt, A., Pinz, A., & Zisserman, A. (2006b). Incremental learning of object detectors using a visual shape alphabet. In Proceedings of the CVPR (Vol. 1, pp. 3\u201310), June 2006.","DOI":"10.1109\/CVPR.2006.153"},{"key":"139_CR43","doi-asserted-by":"crossref","unstructured":"Opelt, A., Pinz, A., & Zisserman, A. (2006c). A boundary-fragment-model for object detection. In Proceedings of the ECCV (Vol. 2, pp. 575\u2013588), May 2006.","DOI":"10.1007\/11744047_44"},{"key":"139_CR44","doi-asserted-by":"crossref","unstructured":"Opelt, A., Pinz, A., & Zisserman, A. (2006d). Fusing shape and appearance information for category detection. In Proceedings of the BMVC (Vol. 1, pp. 117\u2013126), September 2006.","DOI":"10.5244\/C.20.13"},{"issue":"1","key":"139_CR45","doi-asserted-by":"crossref","first-page":"78","DOI":"10.1006\/jecp.2000.2609","volume":"79","author":"P. C. Quinn","year":"2001","unstructured":"Quinn, P. C., Eimas, P. D., & Tarr, M. J. (2001). Perceptual categorization of cat and dog silhouettes by 3-to-4 month old infants. Journal of Experimental Child Psychology, 79(1), 78\u201394.","journal-title":"Journal of Experimental Child Psychology"},{"key":"139_CR46","doi-asserted-by":"crossref","unstructured":"Sali, E., & Ullman, S. (1999). Combining class-specific fragments for object classification. In Proceedings of the BMVC (Vol. 1, pp.\u00a0203\u2013213).","DOI":"10.5244\/C.13.21"},{"key":"139_CR47","doi-asserted-by":"crossref","unstructured":"Seemann, E., Leibe, B., & Schiele, B. (2006). Multi-aspect detection of articulated objects. In Proceedings of the CVPR.","DOI":"10.1109\/CVPR.2006.193"},{"key":"139_CR48","unstructured":"Serre, T., Wolf, L., & Poggio, T. (2005). A new biologically motivated framework for robust object recognition. In Proceedings of the CVPR."},{"key":"139_CR49","doi-asserted-by":"crossref","unstructured":"Shotton, J., Blake, A., & Cipolla, R. (2005). Contour-based learning for object detection. In Proceedings of the ICCV (Vol. 1, pp. 503\u2013510).","DOI":"10.1109\/ICCV.2005.63"},{"key":"139_CR50","doi-asserted-by":"crossref","unstructured":"Shotton, J., Winn, J., Rother, C., & Criminisi, A. (2006). TextonBoost: Joint appearance, shape and context modeling fdor multi-class object recognition and segmentation. In Proceedings of the ECCV (Vol. 1, pp. 1\u201315), May 2006.","DOI":"10.1007\/11744023_1"},{"key":"139_CR51","doi-asserted-by":"crossref","unstructured":"Sivic, J., & Zisserman, A. (2003). Video Google: a text retrieval approach to object matching in videos. In Proceedings of the ICCV.","DOI":"10.1109\/ICCV.2003.1238663"},{"key":"139_CR52","doi-asserted-by":"crossref","unstructured":"Sivic, J., Russell, B., Efros, A., Zisserman, A., & Freeman, W. (2005). Discovering objects and their location in images. In Proceedings of the ICCV.","DOI":"10.1109\/ICCV.2005.77"},{"key":"139_CR53","doi-asserted-by":"crossref","unstructured":"Thomas, A., Ferrari, V., Leibe, B., Tuytelaars, T., Schiele, B., & VanGool, L. (2006). Towards multi-view object detection. In Proceedings of the CVPR.","DOI":"10.1109\/CVPR.2006.311"},{"key":"139_CR54","doi-asserted-by":"crossref","unstructured":"Thureson, J., & Carlsson, S. (2004). Appearance based qualitative image description for object class recognition. In Proceedings of the ECCV (pp. 518\u2013529).","DOI":"10.1007\/978-3-540-24671-8_41"},{"key":"139_CR55","doi-asserted-by":"crossref","unstructured":"Torralba, A., Murphy, K. P., & Freeman, W. T. (2004). Sharing features: efficient boosting procedures for multiclass object detection. In Proceedings of the CVPR.","DOI":"10.1109\/CVPR.2004.1315241"},{"key":"139_CR56","unstructured":"Tu, Z. (2005). Probabilistic boosting-tree: learning discriminative models for classification, recognition, and clustering. In Proceedings of the CVPR (pp. 1589\u20131596)."},{"key":"139_CR57","doi-asserted-by":"crossref","unstructured":"Vidal-Naquet, M., & Ullman, S. (2003). Object recognition with informative features and linear classification. In Proceedings of the ICCV (Vol. 1, pp. 281\u2013288).","DOI":"10.1109\/ICCV.2003.1238356"},{"key":"139_CR58","doi-asserted-by":"crossref","unstructured":"Wang, G., Zhang, Y., & FeiFei, L. (2006). Using dependent regions for object categorization in a generative framework. In Proceedings of the CVPR.","DOI":"10.1109\/CVPR.2006.324"},{"key":"139_CR59","unstructured":"Williams, C. K. I., & Allan, M. (2006). On a connection between object localization with a generative template of features and pose-space prediction methods (Technical Report EDI-INF-RR-0719). School of Informatics, University of Edinburgh."},{"key":"139_CR60","unstructured":"Winn, J., Criminisi, A., & Minka, T. (2005). Object categorization by learning universal visual dictionary. In Proceedings of the ICCV (pp. 1800\u20131807)."},{"key":"139_CR61","unstructured":"Zhang, W., Yu, B., Zelinsky, G. J., & Samaras, D. (2005). Object class recognition using multiple layer boosting with heterogenous features. In Proceedings of the CVPR (pp. 66\u201373)."},{"issue":"2","key":"139_CR62","doi-asserted-by":"crossref","first-page":"213","DOI":"10.1007\/s11263-006-9794-4","volume":"73","author":"J. Zhang","year":"2007","unstructured":"Zhang, J., Marszalek, M., Lazebnik, S., & Schmid, C. (2007). Local features and kernels for classification of texture and object categories: a comprehensive study. International Journal of Computer Vision, 73(2), 213\u2013238.","journal-title":"International Journal of Computer Vision"}],"container-title":["International Journal of Computer Vision"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/link.springer.com\/content\/pdf\/10.1007\/s11263-008-0139-3.pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,1,30]],"date-time":"2025-01-30T06:52:34Z","timestamp":1738219954000},"score":1,"resource":{"primary":{"URL":"http:\/\/link.springer.com\/10.1007\/s11263-008-0139-3"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2008,5,13]]},"references-count":62,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2008,10]]}},"alternative-id":["139"],"URL":"https:\/\/doi.org\/10.1007\/s11263-008-0139-3","relation":{},"ISSN":["0920-5691","1573-1405"],"issn-type":[{"type":"print","value":"0920-5691"},{"type":"electronic","value":"1573-1405"}],"subject":[],"published":{"date-parts":[[2008,5,13]]}}}