Spatial pyramid matching

Svetlana Lazebnik 1 Cordelia Schmid 2 Jean Ponce 3
2 LEAR - Learning and recognition in vision
Inria Grenoble - Rhône-Alpes, LJK - Laboratoire Jean Kuntzmann, INPG - Institut National Polytechnique de Grenoble
Abstract : This chapter deals with the problem of whole-image categorization. We may want to classify a photograph based on a high-level semantic attribute (e.g., indoor or outdoor), scene type (forest, street, office, etc.), or object category (car, face, etc.). Our philosophy is that such global image tasks can be approached in a holistic fashion: It should be possible to develop image representations that use low-level features to directly infer high-level semantic information about the scene without going through the intermediate step of segmenting the image into more "basic" semantic entities. For example, we should be able to recognize that an image contains a beach scene without first segmenting and identifying its separate components, such as sand, water, sky, or bathers. This philosophy is inspired by psychophysical and psychological evidence that people can recognize scenes by considering them in a "holistic" manner, while overlooking most of the details of the constituent objects (Oliva and Torralba, 2001). It has been shown that human subjects can perform high-level categorization tasks extremely rapidly and in the near absence of attention (Thorpe et al., 1996; Fei-Fei et al., 2002), which would most likely preclude any feedback or detailed analysis of individual parts of the scene.
Type de document :
Chapitre d'ouvrage
Sven J. Dickinson and Aleš Leonardis and Bernt Schiele and Michael J. Tarr. Object Categorization: Computer and Human Vision Perspectives, Cambridge University Press, pp.401-415, 2009, 9780521887380
Liste complète des métadonnées

https://hal.inria.fr/inria-00548647
Contributeur : Thoth Team <>
Soumis le : jeudi 6 janvier 2011 - 11:20:59
Dernière modification le : jeudi 29 septembre 2016 - 01:22:39
Document(s) archivé(s) le : jeudi 7 avril 2011 - 02:36:44

Fichier

pyramid_chapter.pdf
Fichiers produits par l'(les) auteur(s)

Identifiants

  • HAL Id : inria-00548647, version 1

Collections

Citation

Svetlana Lazebnik, Cordelia Schmid, Jean Ponce. Spatial pyramid matching. Sven J. Dickinson and Aleš Leonardis and Bernt Schiele and Michael J. Tarr. Object Categorization: Computer and Human Vision Perspectives, Cambridge University Press, pp.401-415, 2009, 9780521887380. <inria-00548647>

Partager

Métriques

Consultations de
la notice

606

Téléchargements du document

742