Search PubMed⌕ Search

PubMed · 427227

Transformation and relational-structure schemes for visual pattern recognition. Two models tested experimentally with rotated random-dot patterns.

Abstract

Two models for visual pattern recognition are described; the one based on application of internal compensatory transformations to pattern representations, the other based on encoding of patterns in terms of local features and spatial relations between these local features. These transformations and relational-structure models are each endowed with the same experimentally observed invariance properties, which include independence to pattern translation and pattern jitter, and, depending on the particular versions of the models, independence to pattern reflection and inversion (180 degrees rotation). Each model is tested by comparing the predicted recognition performance with experimentally determined recognition performance using as stimuli random-dot patterns that were variously rotated in the plane. The level of visual recognition of such patterns is known to depend strongly on rotation angle. It is shown that the relational-structure model equipped with an invariance to pattern inversion gives responses which are in close agreement with the experimental data over all pattern rotation angles. In contrast, the transformation model equipped with the same invariances gives poor agreement to the experimental data. Some implications of these results are considered.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

D H Foster, R J Mason. 1979-03-06. Transformation and relational-structure schemes for visual pattern recognition. Two models tested experimentally with rotated random-dot patterns.. https://doi.org/10.1007/bf00337439

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Representational momentum in perception and grasping: translating versus transforming objects.

Representational momentum is the tendency to misremember the stopping point of a moving object as further forward in the direction of movement. Results of several studies suggest that this effect is typical for changes in position (e.g., translation) and not for changes in object shape (transformation). Additionally, the effect seems to be stronger in motor tasks than in perceptual tasks. Here, participants judged the final distance between two spheres after this distance had been increasing or decreasing. The spheres were two separately translating objects or were connected to form a single transforming object (a dumbbell). Participants also performed a motor task in which they grasped virtual versions of the final objects. We found representational momentum for the visual judgment task for both stimulus types. As predicted, it was stronger for the spheres than for the dumbbells. In contrast, for grasping, only the dumbbells produced representational momentum (larger maximum grip aperture when the dumbbells had been growing compared to when they had been shrinking). Because type of stimulus change had these different effects on representational momentum for perception and action, we conclude that different sources of information are used in the two tasks or that they are governed by different mechanisms.

Field Dependence-Independence↗

Cross-orientation summation in texture segregation.

Human texture vision has been modeled as a filter-rectify-filter (FRF) process, in which '2nd-order' filters detect changes in the rectified outputs of luminance-based '1st-order' filters. This study tested the validity of the two basic assumptions of the standard FRF model, namely (a) that the 2nd-order filters are sensitive to spatial modulations in both contrast and orientation, and (b) that the 2nd-order filters are tuned to different 1st-order orientations. In the first experiment, we tested subthreshold summation between two orthogonal carrier orientations in detection of a texture region, which was defined by contrast modulations across regions in the two carrier orientations, while systematically varying the relative change magnitudes between the two orientations. The results showed that the detection thresholds were determined by spatial difference in the contrast integrated over the two orientations. Orientation difference did act as a segregation cue, but only when there was no differences in carrier contrast. This suggests that two mechanisms are involved in texture segregation; one that detects changes in luminance contrast and another that detects changes in orientation. To further analyze the latter mechanism, a second experiment measured cross-orientation summation in the detection of purely orientation-defined textures, using stimuli that were density modulations of two orientations presented among randomly-orientated distractors. Again, the relative modulation magnitudes between the two orientations was systematically varied. The results are consistent with the notions that (a) the dominant orientation is extracted from the 1st-order outputs before the 2nd-order process, and that (b) the 2nd-order, spatial comparison process integrates those dominant signals over different orientations.

Field Dependence-Independence↗

Context effects on texture border localization bias.

Observers are able to locate precisely a border defined by changes in texture orientation. The prevailing theory is that such localization takes place using a hierarchical, filter-rectify-filter mechanism. An alternative theory is that contextual modulation causes the border elements to stand out. Here we show that perceived border location is inconsistent with contextual modulation from iso-oriented elements. The perceived location of a vertical border defined by vertical texture on one side, and horizontal texture on the other side, is biased towards the vertical texture. We found the same bias in a single row of texture. Therefore, the bias is not due to contextual influences from surrounding iso-oriented elements. Contextual influences between cross-oriented elements can explain the data.

Field Dependence-Independence↗