A Statistical Model for General Contextual Object Recognition

Carbonetto, P; Freitas, N; Barnard, K

Conference item

A Statistical Model for General Contextual Object Recognition

Abstract:: We consider object recognition as the process of attaching meaningful labels to specific regions of an image, and propose a model that learns spatial relationships between objects. Given a set of images and their associated text (e.g. keywords, captions, descriptions), the objective is to segment an image, in either a crude or sophisticated fashion, then to find the proper associations between words and regions. Previous models are limited by the scope of the representation. In particular, they fail to exploit spatial context in the images and words. We develop a more expressive model that takes this into account. We formulate a spatially consistent probabilistic mapping between continuous image feature vectors and the supplied word tokens. By learning both word-to-region associations and object relations, the proposed model augments scene segmentations due to smoothing implicit in spatial consistency. Context introduces cycles to the undirected graph, so we cannot rely on a straightforward implementation of the EM algorithm for estimating the model parameters and densities of the unknown alignment variables. Instead, we develop an approximate EM algorithm that uses loopy belief propagation in the inference step and iterative scaling on the pseudo-likelihood approximation in the parameter update step. The experiments indicate that our approximate inference and learning algorithm converges to good local solutions. Experiments on a diverse array of images show that spatial context considerably improves the accuracy of object recognition. Most significantly, spatial context combined with a nonlinear discrete object representation allows our models to cope well with over-segmented scenes.

Actions

Email

Email this record

Send the bibliographic details of this record to your email address.

Your Email
Please enter the email address that the record information will be sent to.

-
Your message (optional)
Please add any additional information to be included within the email.
Cite

Cite this record

APA Style

Carbonetto, P., Freitas, N., & Barnard, K. (2004). A Statistical Model for General Contextual Object Recognition. European Conference on Computer Vision (ECCV), 3021.

MLA Style

Carbonetto, P., et al. “A Statistical Model for General Contextual Object Recognition.” European Conference on Computer Vision (ECCV), vol. 3021, Springer Berlin Heidelberg, 2004.

Chicago Style

Carbonetto, P, N Freitas, and K Barnard. 2004. “A Statistical Model for General Contextual Object Recognition.” In European Conference on Computer Vision (ECCV). Vol. 3021. Springer Berlin Heidelberg.
Share
Print