Search PubMed⌕ Search

PubMed · 17153187

Object recognition in dense clutter.

Abstract

Observers in recognition experiments invariably view objects against a blank background, whereas observers of real scenes sometimes view objects against dense clutter. In this study, we examined whether an object's background affects the information used for recognition. Our stimuli consisted of color photographs of everyday objects. The photographs were organized either as a sparse array, as is typical of a visual search experiment, or as high density clutter, such as might be found in a toy chest, a handbag, or a kitchen drawer. The observer's task was to locate an animal, vehicle, or food target in the stimulus. We varied the information in the stimuli by convolving them with a low-pass filter (blur) or a high-pass filter (edge) or converting them to grayscale. In two experiments, we found that the blur and edge manipulations produced a modest decrement in performance with the sparse arrangement but a severe decrement in performance with the clutter arrangement. These results indicate that the information used for recognition depends on the object's background. Thus, models of recognition that have been developed for isolated objects may not generalize to objects in dense clutter.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Mary J Bravo, Hany Farid. 2006. Object recognition in dense clutter.. https://doi.org/10.3758/bf03193354

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Absence of flash-lag when judging global shape from local positions.

When a flash is presented aligned with a moving stimulus, the former is perceived to lag behind the latter (the flash-lag effect). We study whether this mislocalization occurs when a positional judgment is not required, but a veridical spatial relationship between moving and flashed stimuli is needed to perceive a global shape. To do this, we used Glass patterns that are formed by pairs of correlated dots. One dot of each pair was presented moving and, at a given moment, the other dot of each pair was flashed in order to build the Glass pattern. If a flash-lag effect occurs between each pair of dots, we expect the best perception of the global shape to occur when the flashed dots are presented before the moving dots arrive at the position that physically builds the Glass pattern. Contrary to this, we found that the best detection of Glass patterns occurred for the situation of physical alignment. This result is not consistent with a low-level contribution to the flash-lag effect.

Form Perception↗

Object recognition and segmentation by a fragment-based hierarchy.

How do we learn to recognize visual categories, such as dogs and cats? Somehow, the brain uses limited variable examples to extract the essential characteristics of new visual categories. Here, I describe an approach to category learning and recognition that is based on recent computational advances. In this approach, objects are represented by a hierarchy of fragments that are extracted during learning from observed examples. The fragments are class-specific features and are selected to deliver a high amount of information for categorization. The same fragments hierarchy is then used for general categorization, individual object recognition and object-parts identification. Recognition is also combined with object segmentation, using stored fragments, to provide a top-down process that delineates object boundaries in complex cluttered scenes. The approach is computationally effective and provides a possible framework for categorization, recognition and segmentation in human vision.

Form Perception↗

Dynamics of shape interaction in human vision.

Spatial context can alter perceived shape, and temporal context can influence the perception of a stimulus. We sought to determine the time course of shape interactions by using a paradigm in which closed shape contours are laterally displaced over space and time. Target and masks are separated by various stimulus onset asynchrony (SOA) values, yielding forward, backward, and simultaneous masking conditions. Results indicate that spatial lateral interactions of shape are amplified by temporal asynchrony, reaching a peak at SOAs of 80-110 ms. Mask amplitude scales all effects and masking is shape specific. When a single mask follows the target, both spatial configuration and mask onset transient are critical in determining depth of masking. When the target is followed by two sequential masks, the possibility of apparent motion determines whether one or both masks drive masking. These findings suggest that temporal interactions of shape are dependent on an interactive combination of shape specificity and transients, that apparent motion plays a modulatory role, and that target shape is determined after a temporal window, not at its onset.

Form Perception↗