Search PubMed⌕ Search

PubMed · 8935905

Structure from motion: a tolerance analysis.

Abstract

We present a tolerance analysis that is applicable to a large group of stimuli used in structure-from-motion tasks. Human performance in structure-from-motion tasks reflects the fact that the visual system deals with projections of a 3-D world on the retina. A tolerance analysis reveals the relationship between the projections and the 3-D world. Any realistic model of the visual system should incorporate a tolerance analysis as a complete description of the stimulus. By way of example we apply the tolerance analysis to the stimuli used in two widely known experiments in which different properties of structure were tested--that is, perceived nonrigidity (Norman & Todd, 1993) and ordering in depth (Hildreth, Grzywacz, Adelson, & Inada, 1990). The analysis explains qualitatively the results of these experiments, illustrating that the results are to a large extent due to stimulus limitations rather than to mechanistic properties of the visual system. From our analysis it follows that far more sensitive measurements of the optic information are needed to obtain metric structure than affine structure.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

M A Hogervorst, A M Kappers, J J Koenderink. 1996. Structure from motion: a tolerance analysis.. https://doi.org/10.3758/bf03206820

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Miniaturized three-dimensional endoscopic imaging system based on active stereovision.

A miniaturized three-dimensional endoscopic imaging system is presented. The system consists of two imaging in channels that can be used to obtain an image from an object of interest and to project as tructured light onto the imaged object to measure the surface topology. The structured light was generated with a collimated monochromatic light source and a holographic binary phase grating. The imaging and projection channels were calibrated by use of a modified pinhole camera. The surface profile was extracted by use of triangulation between the projected feature points and the two channel ofthe endoscope. The imaging system was evaluated in three-dimensional measurements of several objects with known geometries. The results show that surface profiles of the objects with different surfaces and dimensions can be obtained at high accuracy. The in vivo measurements at tissue sites of human skin and an oral cavity demonstrated the potential of the technique for clinical applications.

Depth Perception↗

Computer-enhanced stereoscopic vision in a head-mounted operating binocular.

Based on the Varioscope, a commercially available head-mounted operating binocular, we have developed the Varioscope AR, a see through head-mounted display (HMD) for augmented reality visualization that seamlessly fits into the infrastructure of a surgical navigation system. We have assessed the extent to which stereoscopic visualization improves target localization in computer-aided surgery in a phantom study. In order to quantify the depth perception of a user aiming at a given target, we have designed a phantom simulating typical clinical situations in skull base surgery. Sixteen steel spheres were fixed at the base of a bony skull, and several typical craniotomies were applied. After having taken CT scans, the skull was filled with opaque jelly in order to simulate brain tissue. The positions of the spheres were registered using VISIT, a system for computer-aided surgical navigation. Then attempts were made to locate the steel spheres with a bayonet probe through the craniotomies using VISIT and the Varioscope AR as a stereoscopic display device. Localization of targets 4 mm in diameter using stereoscopic vision and additional visual cues indicating target proximity had a success rate (defined as a first-trial hit rate) of 87.5%. Using monoscopic vision and target proximity indication, the success rate was found to be 66.6%. Omission of visual hints on reaching a target yielded a success rate of 79.2% in the stereo case and 56.25% with monoscopic vision. Time requirements for localizing all 16 targets ranged from 7.5 min (stereo, with proximity cues) to 10 min (mono, without proximity cues). Navigation error is primarily governed by the accuracy of registration in the navigation system, whereas the HMD does not appear to influence localization significantly. We conclude that stereo vision is a valuable tool in augmented reality guided interventions.

Depth Perception↗

Is neural filling-in necessary to explain the perceptual completion of motion and depth information?

Retinal activity is the first stage of visual perception. Retinal sampling is non-uniform and not continuous, yet visual experience is not characterized by holes and discontinuities in the world. How does the brain achieve this perceptual completion? Fifty years ago, it was suggested that visual perception involves a two-stage process of (i) edge detection followed by (ii) neural filling-in of surface properties. We examine whether this general hypothesis can account for the specific example of perceptual completion of a small target surrounded by dynamic dots (an 'artificial scotoma'), a phenomenon argued to provide insight into the mechanisms responsible for perception. We degrade the target's borders using first blur and then depth continuity, and find that border degradation does not influence time to target disappearance. This indicates that important information for the continuity of target perception is conveyed at a coarse spatial scale. We suggest that target disappearance could result from adaptation that is not specific to borders, and question the need to hypothesize an active filling-in process to explain this phenomenon.

Depth Perception↗