Search PubMed⌕ Search

PubMed · 1502274

Dynamic binding in a neural network for shape recognition.

Abstract

Given a single view of an object, humans can readily recognize that object from other views that preserve the parts in the original view. Empirical evidence suggests that this capacity reflects the activation of a viewpoint-invariant structural description specifying the object's parts and the relations among them. This article presents a neural network that generates such a description. Structural description is made possible through a solution to the dynamic binding problem: Temporary conjunctions of attributes (parts and relations) are represented by synchronized oscillatory activity among independent units representing those attributes. Specifically, the model uses synchrony (a) to parse images into their constituent parts, (b) to bind together the attributes of a part, and (c) to bind the relations to the parts to which they apply. Because it conjoins independent units temporarily, dynamic binding allows tremendous economy of representation and permits the representation to reflect the attribute structure of the shapes represented.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

J E Hummel, I Biederman. 1992. Dynamic binding in a neural network for shape recognition.. https://doi.org/10.1037/0033-295x.99.3.480

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

The effects of phase on the perception of 3D shape from texture: psychophysics and modeling.

Two experiments are reported in which observers judged the apparent shapes of elliptical cylinders with eight different textures that were presented with scrambled and unscrambled phase spectra. The results revealed that the apparent depths of these surfaces varied linearly with the ground truth in all conditions, and that the overall magnitude of surface relief was systematically underestimated. In general, the apparent depth of a surface is significantly attenuated when the phase spectrum of its texture is randomly scrambled, though the magnitude of this effect varies for different types of texture. A new computational model of 3D shape from texture is proposed in which apparent depth is estimated from the relative density of edges in different local regions of an image, and the predictions of this model are highly correlated with the observers' judgments.

Depth Perception↗

Stereomotion suppression and the perception of speed: accuracy and precision as a function of 3D trajectory.

The precision and accuracy of speed discrimination performance for stereomotion stimuli were assessed for several receding 3D trajectories confined to the horizontal meridian. It has previously been demonstrated in a variety of tasks that detection thresholds are substantially higher when subjects observe a stereomotion stimulus than when simply viewing one of its component monocular half-images--a phenomenon known as stereomotion suppression (C. W. Tyler, 1971). Using monocularly visible motion in depth targets, we found mean speed discrimination thresholds to be higher for stereomotion, compared with monocular lateral speed discrimination thresholds for equivalent stimuli, demonstrating a disadvantage for binocular viewing in the case of speed discrimination as well. Furthermore, speed discrimination thresholds for motion in depth were not systematically affected by trajectory angle; hence, the disadvantage of binocular viewing persists even when there are concurrent changes in binocular visual direction. Lastly, there was a tendency for oblique trajectories of stereomotion to be perceived as faster than equally rapid motion receding directly away from the subject along the midline. Our data, in addition to earlier stereomotion suppression observations, are consistent with a stereomotion system that takes a noisy, weighted difference of the stimulus velocities in the two eyes to compute motion in depth.

Depth Perception↗

Visual saliency and texture segregation without feature gradient.

A central notion in the study of texture segregation is that of feature gradient (or feature contrast). In orientation-based texture segregation, orientation gradients have indeed played a fundamental role in explaining behavioral results. Here, however, we show that general, smoothly varying, orientation-defined textures (ODTs) exhibit striking perceptual singularities that are completely unpredictable from orientation gradients. These singularities defy not only popular texture segregation theories but also virtually all computational segmentation methods, and they confound previous behavioral studies with smoothly varying ODTs. We provide psychophysical evidence that perceptual singularities in smooth ODTs are salient visual features consistent across observers and with significant effect on the perception and segregation of oriented textures. We further show that, although orientation gradients cannot predict them, perceptual singularities in smooth ODTs emerge directly from, and can be spatially localized by, two ODT curvatures. Given the traditional role of feature gradients in early vision, the significance of these findings extends well beyond orientation-based texture segregation to issues ranging from curve integration and fragment grouping, through the perception of 3D shape, to the functional organization of the primary visual cortex.

Depth Perception↗