Search NASA⌕ Search

SEARCH · Search NASA

Results for “Visual Perception”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 307 records · Page 17

Manual control of yaw motion with combined visual and vestibular cues

Measurements are made of manual control performance in the closed-loop task of nulling perceived self-rotation velocity about an earth-vertical axis. Self-velocity estimation was modelled as a function of the simultaneous presentation of vestibular and peripheral visual field motion cues. Based on measured low-frequency operator behavior in three visual field environments, a parallel channel linear model is proposed which has separate visual and vestibular pathways summing in a complementary manner. A correction to the frequency responses is provided by a separate measurement of manual control performance in an analogous visual pursuit nulling task. The resulting dual-input describing function for motion perception dependence on combined cue presentation supports the complementary model, in which vestibular cues dominate sensation at frequencies above 0.05 Hz. The describing function model is extended by the proposal of a non-linear cue conflict model, in which cue weighting depends on the level of agreement between visual and vestibular cues.

Zacharias, G. L.↗

SAVA 3: A testbed for integration and control of visual processes

The development of an experimental test-bed to investigate the integration and control of perception in a continuously operating vision system is described. The test-bed integrates a 12 axis robotic stereo camera head mounted on a mobile robot, dedicated computer boards for real-time image acquisition and processing, and a distributed system for image description. The architecture was designed to: (1) be continuously operating, (2) integrate software contributions from geographically dispersed laboratories, (3) integrate description of the environment with 2D measurements, 3D models, and recognition of objects, (4) capable of supporting diverse experiments in gaze control, visual servoing, navigation, and object surveillance, and (5) dynamically reconfiguarable.

Crowley, James L.↗

Sensing Super-position: Visual Instrument Sensor Replacement

The coming decade of fast, cheap and miniaturized electronics and sensory devices opens new pathways for the development of sophisticated equipment to overcome limitations of the human senses. This project addresses the technical feasibility of augmenting human vision through Sensing Super-position using a Visual Instrument Sensory Organ Replacement (VISOR). The current implementation of the VISOR device translates visual and other passive or active sensory instruments into sounds, which become relevant when the visual resolution is insufficient for very difficult and particular sensing tasks. A successful Sensing Super-position meets many human and pilot vehicle system requirements. The system can be further developed into cheap, portable, and low power taking into account the limited capabilities of the human user as well as the typical characteristics of his dynamic environment. The system operates in real time, giving the desired information for the particular augmented sensing tasks. The Sensing Super-position device increases the image resolution perception and is obtained via an auditory representation as well as the visual representation. Auditory mapping is performed to distribute an image in time. The three-dimensional spatial brightness and multi-spectral maps of a sensed image are processed using real-time image processing techniques (e.g. histogram normalization) and transformed into a two-dimensional map of an audio signal as a function of frequency and time. This paper details the approach of developing Sensing Super-position systems as a way to augment the human vision system by exploiting the capabilities of the human hearing system as an additional neural input. The human hearing system is capable of learning to process and interpret extremely complicated and rapidly changing auditory patterns. The known capabilities of the human hearing system to learn and understand complicated auditory patterns provided the basic motivation for developing an image-to-sound mapping system.

Maluf, David A.↗

Sensing Super-Position: Human Sensing Beyond the Visual Spectrum

The coming decade of fast, cheap and miniaturized electronics and sensory devices opens new pathways for the development of sophisticated equipment to overcome limitations of the human senses. This paper addresses the technical feasibility of augmenting human vision through Sensing Super-position by mixing natural Human sensing. The current implementation of the device translates visual and other passive or active sensory instruments into sounds, which become relevant when the visual resolution is insufficient for very difficult and particular sensing tasks. A successful Sensing Super-position meets many human and pilot vehicle system requirements. The system can be further developed into cheap, portable, and low power taking into account the limited capabilities of the human user as well as the typical characteristics of his dynamic environment. The system operates in real time, giving the desired information for the particular augmented sensing tasks. The Sensing Super-position device increases the image resolution perception and is obtained via an auditory representation as well as the visual representation. Auditory mapping is performed to distribute an image in time. The three-dimensional spatial brightness and multi-spectral maps of a sensed image are processed using real-time image processing techniques (e.g. histogram normalization) and transformed into a two-dimensional map of an audio signal as a function of frequency and time. This paper details the approach of developing Sensing Super-position systems as a way to augment the human vision system by exploiting the capabilities of Lie human hearing system as an additional neural input. The human hearing system is capable of learning to process and interpret extremely complicated and rapidly changing auditory patterns. The known capabilities of the human hearing system to learn and understand complicated auditory patterns provided the basic motivation for developing an image-to-sound mapping system. The human brain is superior to most existing computer systems in rapidly extracting relevant information from blurred, noisy, and redundant images. From a theoretical viewpoint, this means that the available bandwidth is not exploited in an optimal way. While image-processing techniques can manipulate, condense and focus the information (e.g., Fourier Transforms), keeping the mapping as direct and simple as possible might also reduce the risk of accidentally filtering out important clues. After all, especially a perfect non-redundant sound representation is prone to loss of relevant information in the non-perfect human hearing system. Also, a complicated non-redundant image-to-sound mapping may well be far more difficult to learn and comprehend than a straightforward mapping, while the mapping system would increase in complexity and cost. This work will demonstrate some basic information processing for optimal information capture for headmounted systems.

Maluf, David A.↗

A Reevaluation of the Vestibulo-Ocular Reflex: New Ideas of its Purpose, Properties, Neural Substrate, and Disorders

Conventional views of the Vestibulo-Ocular Reflex (VOR) have emphasized testing with caloric stimuli and by passively rotating patients at low frequencies in a chair. The properties of the VOR tested under these conditions differ from the performance of this reflex during the natural function for which it evolved-locomotion. Only the VOR (and not visually mediated eye movements) can cope with the high-frequency angular and linear perturbations of the head that occur during locomotion; this is achieved by generating eye movements at short latency (less than 16 msec). Interpretation of vestibular testing is enhanced by the realization that, although the di- and trisynaptic components of the VOR are essential for this short-latency response, the overall accuracy and plasticity of the VOR depend upon a distributed, parallel network of neurons involving the vestibular nuclei. Neurons in this network variously encode inputs from the labyrinthine semicircular canals and otoliths, as well as from the visual and somatosensory systems. The central vestibular pathways branch to contact vestibular cortex (for perception) and the spinal cord (for control of posture). Thus, the vestibular nuclei basically coordinate the stabilization of gaze and posture, and contribute to the perception of verticality and self-motion. Consequently, brainstem disorders that disrupt the VOR cause not just only nystagmus, but also instability of posture (eg, increased fore-aft sway in patients with downbeat nystagmus) and disturbance of spatial orientation (eg, tilt of the subjective visual vertical in Wallenberg's syndrome).

Leigh, R. John↗

A reevaluation of the vestibulo-ocular reflex: new ideas of its purpose, properties, neural substrate, and disorders

Conventional views of the vestibulo-ocular reflex (VOR) have emphasized testing with caloric stimuli and by passively rotating patients at low frequencies in a chair. The properties of the VOR tested under these conditions differ from the performance of this reflex during the natural function for which it evolved--locomotion. Only the VOR (and not visually mediated eye movements) can cope with the high-frequency angular and linear perturbations of the head that occur during locomotion; this is achieved by generating eye movements at short latency (< 16 msec). Interpretation of vestibular testing is enhanced by the realization that, although the di- and trisynaptic components of the VOR are essential for this short-latency response, the overall accuracy and plasticity of the VOR depend upon a distributed, parallel network of neurons involving the vestibular nuclei. Neurons in this network variously upon a distributed, parallel network of neurons involving the vestibular nuclei. Neurons in this network variously encode inputs from the labyrinthine semicircular canals and otoliths, as well as from the visual and somatosensory systems. The central vestibular pathways branch to contact vestibular cortex (for perception) and the spinal cord (for control of posture). Thus, the vestibular nuclei basically coordinate the stabilization of gaze and posture, and contribute to the perception of verticality and self-motion. Consequently, brainstem disorders that disrupt the VOR cause not just only nystagmus, but also instability of posture (eg, increased fore-aft sway in patients with downbeat nystagmus) and disturbance of spatial orientation (eg, tilt of the subjective visual vertical in Wallenberg's syndrome).

Review↗

Image gathering and coding for digital restoration: Information efficiency and visual quality

Image gathering and coding are commonly treated as tasks separate from each other and from the digital processing used to restore and enhance the images. The goal is to develop a method that allows us to assess quantitatively the combined performance of image gathering and coding for the digital restoration of images with high visual quality. Digital restoration is often interactive because visual quality depends on perceptual rather than mathematical considerations, and these considerations vary with the target, the application, and the observer. The approach is based on the theoretical treatment of image gathering as a communication channel (J. Opt. Soc. Am. A2, 1644(1985);5,285(1988). Initial results suggest that the practical upper limit of the information contained in the acquired image data range typically from approximately 2 to 4 binary information units (bifs) per sample, depending on the design of the image-gathering system. The associated information efficiency of the transmitted data (i.e., the ratio of information over data) ranges typically from approximately 0.3 to 0.5 bif per bit without coding to approximately 0.5 to 0.9 bif per bit with lossless predictive compression and Huffman coding. The visual quality that can be attained with interactive image restoration improves perceptibly as the available information increases to approximately 3 bifs per sample. However, the perceptual improvements that can be attained with further increases in information are very subtle and depend on the target and the desired enhancement.

Huck, Friedrich O.↗

Real-Time Cognitive Computing Architecture for Data Fusion in a Dynamic Environment

A novel cognitive computing architecture is conceptualized for processing multiple channels of multi-modal sensory data streams simultaneously, and fusing the information in real time to generate intelligent reaction sequences. This unique architecture is capable of assimilating parallel data streams that could be analog, digital, synchronous/asynchronous, and could be programmed to act as a knowledge synthesizer and/or an "intelligent perception" processor. In this architecture, the bio-inspired models of visual pathway and olfactory receptor processing are combined as processing components, to achieve the composite function of "searching for a source of food while avoiding the predator." The architecture is particularly suited for scene analysis from visual data and odorant.

Duong, Tuan A.↗

Augmentation of Cognition and Perception Through Advanced Synthetic Vision Technology

Synthetic Vision System technology augments reality and creates a virtual visual meteorological condition that extends a pilot's cognitive and perceptual capabilities during flight operations when outside visibility is restricted. The paper describes the NASA Synthetic Vision System for commercial aviation with an emphasis on how the technology achieves Augmented Cognition objectives.

Prinzel, Lawrence J., III↗

Visual-vestibular integration as a function of adaptation to space flight and return to Earth

Research on perception and control of self-orientation and self-motion addresses interactions between action and perception . Self-orientation and self-motion, and the perception of that orientation and motion are required for and modified by goal-directed action. Detailed Supplementary Objective (DSO) 604 Operational Investigation-3 (OI-3) was designed to investigate the integrated coordination of head and eye movements within a structured environment where perception could modify responses and where response could be compensatory for perception. A full understanding of this coordination required definition of spatial orientation models for the microgravity environment encountered during spaceflight.

Reschke, Millard R.↗

The Barberplaid Illusion

Mulligan showed that the perceived direction of a moving grating can be biased by the shape of the Gaussian window in which it is viewed. We sought to determine if a 2-D pattern with an unambiguous velocity would also show such biases. Observers viewed a drifting plaid (sum of two orthogonal 2.5 c/d sinusoidal gratings of 12% contrast, each with a TF of 4 Hz.) whose contrast was modulated spatially by a stationary, asymmetric 2-D Gaussian window (i.e. unequal standard deviations in the principal directions). The direction of plaid motion with respect to the orientation of the window's major axis (Delta Theta) was varied while all other motion parameters were held fixed. Observers reported the perceived plaid direction of motion by adjusting the orientation of a pointer. All five observers showed systematic biases in perceived plaid direction that depended on Delta Theta and the aspect ratio of the Gaussian window (lambda). For circular Gaussian windows Lambda = 1), plaid direction was veridically perceived. However, biases of up to 10 deg. were found for lambda = 2 and Delta Theta = 30 deg. These data present a challenge to models of motion perception which do not explicitly consider the integration of information across the visual field.

Beutter, B. R.↗

Simultaneous dual-task performance reveals parallel response selection after practice

E. H. Schumacher, T. L. Seymour, J. M. Glass, D. E. Kieras, and D. E. Meyer (2001) reported that dual-task costs are minimal when participants are practiced and give the 2 tasks equal emphasis. The present research examined whether such findings are compatible with the operation of an efficient response selection bottleneck. Participants trained until they were able to perform both tasks simultaneously without interference. Novel stimulus pairs produced no reaction time costs, arguing against the development of compound stimulus-response associations (Experiment 1). Manipulating the relative onsets (Experiments 2 and 4) and durations (Experiments 3 and 4) of response selection processes did not lead to dual-task costs. The results indicate that the 2 tasks did not share a bottleneck after practice.

Practice (Psychology)↗

The relationship between perceived length and egocentric location in Muller-Lyer figures with one versus two chevrons

We examined the apparent dissociation of perceived length and perceived position with respect to the Muller-Lyer (M-L) illusion. With the traditional (two-chevron) figure, participants made accurate open-loop pointing responses at the endpoints of the shaft, despite the presence of a strong length illusion. This apparently non-Euclidean outcome replicated that of Mack, Heuer, Villardi, and Chambers (1985) and Gillam and Chambers (1985) and contradicts any theory of the M-L illusion in which mislocalization of shaft endpoints plays a role. However, when one of the chevrons was removed, a constant pointing error occurred in the predicted direction, as well as a strong length illusion. Thus, with one-chevron stimuli, perceived length and location were no longer completely dissociated. We speculated that the presence of two opposing chevrons suppresses the mislocalizing effects of a single chevron, especially for figures with relatively short shafts.

NASA Center ARC↗

Motion coherence affects human perception and pursuit similarly

Pursuit and perception both require accurate information about the motion of objects. Recovering the motion of objects by integrating the motion of their components is a difficult visual task. Successful integration produces coherent global object motion, while a failure to integrate leaves the incoherent local motions of the components unlinked. We compared the ability of perception and pursuit to perform motion integration by measuring direction judgments and the concomitant eye-movement responses to line-figure parallelograms moving behind stationary rectangular apertures. The apertures were constructed such that only the line segments corresponding to the parallelogram's sides were visible; thus, recovering global motion required the integration of the local segment motion. We investigated several potential motion-integration rules by using stimuli with different object, vector-average, and line-segment terminator-motion directions. We used an oculometric decision rule to directly compare direction discrimination for pursuit and perception. For visible apertures, the percept was a coherent object, and both the pursuit and perceptual performance were close to the object-motion prediction. For invisible apertures, the percept was incoherently moving segments, and both the pursuit and perceptual performance were close to the terminator-motion prediction. Furthermore, both psychometric and oculometric direction thresholds were much higher for invisible apertures than for visible apertures. We constructed a model in which both perception and pursuit are driven by a shared motion-processing stage, with perception having an additional input from an independent static-processing stage. Model simulations were consistent with our perceptual and oculomotor data. Based on these results, we propose the use of pursuit as an objective and continuous measure of perceptual coherence. Our results support the view that pursuit and perception share a common motion-integration stage, perhaps within areas MT or MST.

NASA Center ARC↗

Simulation evaluation of the effects of time delay and motion on rotorcraft handling qualities

A study aimed at determining the effects of simulator characteristics on perceived handling qualities is discussed. Evaluations were conducted with a baseline set of rotorcraft dynamics, using a simple transfer-function model of an uncoupled helicopter, under different conditions of visual and overall time delays. As the visual and motion parameters were changed, differences in pilot opinion were found reflecting a change in the pilots' perceptions of handling qualities, rather than changes in the aircraft model itself. It is concluded that it is necessary to tailor the motion washout dynamics to suit the task, with reduced washouts for precision maneuvering as compared to aggressive maneuvering. Visual-delay data suggest that it may be better to allow some time delay in the visual path to minimize the mismatch between visual and motion, rather than eliminate the visual delay entirely through lead compensation.

Mitchell, David G.↗

Modeling Visual, Vestibular and Oculomotor Interactions in Self-Motion Estimation

A computational model of human self-motion perception has been developed in collaboration with Dr. Leland S. Stone at NASA Ames Research Center. The research included in the grant proposal sought to extend the utility of this model so that it could be used for explaining and predicting human performance in a greater variety of aerospace applications. This extension has been achieved along with physiological validation of the basic operation of the model.

Perrone, John↗