Search NASA⌕ Search

SEARCH · Search NASA

Results for “Perceptual Masking”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

36 records · Page 2

Perceptual compression of magnitude-detected synthetic aperture radar imagery

A perceptually-based approach for compressing synthetic aperture radar (SAR) imagery is presented. Key components of the approach are a multiresolution wavelet transform, a bit allocation mask based on an empirical human visual system (HVS) model, and hybrid scalar/vector quantization. Specifically, wavelet shrinkage techniques are used to segregate wavelet transform coefficients into three components: local means, edges, and texture. Each of these three components is then quantized separately according to a perceptually-based bit allocation scheme. Wavelet coefficients associated with local means and edges are quantized using high-rate scalar quantization while texture information is quantized using low-rate vector quantization. The impact of the perceptually-based multiresolution compression algorithm on visual image quality, impulse response, and texture properties is assessed for fine-resolution magnitude-detected SAR imagery; excellent image quality is found at bit rates at or above 1 bpp along with graceful performance degradation at rates below 1 bpp.

Gorman, John D.↗

Image-adapted visually weighted quantization matrices for digital image compression

A method for performing image compression that eliminates redundant and invisible image components is presented. The image compression uses a Discrete Cosine Transform (DCT) and each DCT coefficient yielded by the transform is quantized by an entry in a quantization matrix which determines the perceived image quality and the bit rate of the image being compressed. The present invention adapts or customizes the quantization matrix to the image being compressed. The quantization matrix comprises visual masking by luminance and contrast techniques and by an error pooling technique all resulting in a minimum perceptual error for any given bit rate, or minimum bit rate for a given perceptual error.

Watson, Andrew B.↗

A visual detection model for DCT coefficient quantization

The discrete cosine transform (DCT) is widely used in image compression and is part of the JPEG and MPEG compression standards. The degree of compression and the amount of distortion in the decompressed image are controlled by the quantization of the transform coefficients. The standards do not specify how the DCT coefficients should be quantized. One approach is to set the quantization level for each coefficient so that the quantization error is near the threshold of visibility. Results from previous work are combined to form the current best detection model for DCT coefficient quantization noise. This model predicts sensitivity as a function of display parameters, enabling quantization matrices to be designed for display situations varying in luminance, veiling light, and spatial frequency related conditions (pixel size, viewing distance, and aspect ratio). It also allows arbitrary color space directions for the representation of color. A model-based method of optimizing the quantization matrix for an individual image was developed. The model described above provides visual thresholds for each DCT frequency. These thresholds are adjusted within each block for visual light adaptation and contrast masking. For given quantization matrix, the DCT quantization errors are scaled by the adjusted thresholds to yield perceptual errors. These errors are pooled nonlinearly over the image to yield total perceptual error. With this model one may estimate the quantization matrix for a particular image that yields minimum bit rate for a given total perceptual error, or minimum perceptual error for a given bit rate. Custom matrices for a number of images show clear improvement over image-independent matrices. Custom matrices are compatible with the JPEG standard, which requires transmission of the quantization matrix.

Ahumada, Albert J., Jr.↗

The Accuracy of Saccadic and Perceptual Decisions in Visual Search

Saccadic eye movements during search for a target embedded in noise are suboptimally guided by information about target location. Our goal is to compare the spatial information used to guide the saccades with that used for the perceptual decision. Three observers were asked to determine the location of a bright disk (diameter = 21 min) in white noise (signal-to-noise ratio = 4.2) from among 10 possible locations evenly spaced at 5.9 deg eccentricity. In the first of four conditions, observers used natural eye movements. In the three remaining conditions, observers fixated a central cross at all times. The fixation conditions consisted of three different presentation times (100, 200, 300 msec), each followed by a mask. Eye-position data were collected, with a resolution of (approximately) 0.2 deg. In the natural viewing condition, we measured. the accuracy with respect to the target and the latency of the first saccade. In the fixation conditions, we discarded trials in which observers broke fixation. Perceptual performance was computed for all conditions. Averaged across observers, the first saccade was correct (closest to the target location) for 56 +/- (SD) % of trials (chance = 10 %) and occurred after a latency of 313 +/- 56 msec. Perceptual performance averaged 53 +/- 4, 63 +/- 4, 65 +/- 2 % correct at 100, 200, and 300 msec, respectively. For the signal-to-noise ratio used, at the time of initiation of the first saccade, there is little difference between the amount of information about target location available to the perceptual and saccadic systems.

Eckstein, Miguel P.↗

DCT quantization matrices visually optimized for individual images

This presentation describes how a vision model incorporating contrast sensitivity, contrast masking, and light adaptation is used to design visually optimal quantization matrices for Discrete Cosine Transform image compression. The Discrete Cosine Transform (DCT) underlies several image compression standards (JPEG, MPEG, H.261). The DCT is applied to 8x8 pixel blocks, and the resulting coefficients are quantized by division and rounding. The 8x8 'quantization matrix' of divisors determines the visual quality of the reconstructed image; the design of this matrix is left to the user. Since each DCT coefficient corresponds to a particular spatial frequency in a particular image region, each quantization error consists of a local increment or decrement in a particular frequency. After adjustments for contrast sensitivity, local light adaptation, and local contrast masking, this coefficient error can be converted to a just-noticeable-difference (jnd). The jnd's for different frequencies and image blocks can be pooled to yield a global perceptual error metric. With this metric, we can compute for each image the quantization matrix that minimizes the bit-rate for a given perceptual error, or perceptual error for a given bit-rate. Implementation of this system demonstrates its advantages over existing techniques. A unique feature of this scheme is that the quantization matrix is optimized for each individual image. This is compatible with the JPEG standard, which requires transmission of the quantization matrix.

Watson, Andrew B.↗

Image discrimination models predict detection in fixed but not random noise

By means of a two-interval forced-choice procedure, contrast detection thresholds for an aircraft positioned on a simulated airport runway scene were measured with fixed and random white-noise masks. The term fixed noise refers to a constant, or unchanging, noise pattern for each stimulus presentation. The random noise was either the same or different in the two intervals. Contrary to simple image discrimination model predictions, the same random noise condition produced greater masking than the fixed noise. This suggests that observers seem unable to hold a new noisy image for comparison. Also, performance appeared limited by internal process variability rather than by external noise variability, since similar masking was obtained for both random noise types.

NASA Center ARC↗

Tactile Cueing as a Gravitational Substitute for Spatial Navigation During Parabolic Flight

INTRODUCTION: Spatial navigation requires an accurate awareness of orientation in your environment. The purpose of this experiment was to examine how spatial awareness was impaired with changing gravitational cues during parabolic flight, and the extent to which vibrotactile feedback of orientation could be used to help improve performance. METHODS: Six subjects were restrained in a chair tilted relative to the plane floor, and placed at random positions during the start of the microgravity phase. Subjects reported their orientation using verbal reports, and used a hand-held controller to point to a desired target location presented using a virtual reality video mask. This task was repeated with and without constant tactile cueing of "down" direction using a belt of 8 tactors placed around the mid-torso. Control measures were obtained during ground testing using both upright and tilted conditions. RESULTS: Perceptual estimates of orientation and pointing accuracy were impaired during microgravity or during rotation about an upright axis in 1g. The amount of error was proportional to the amount of chair displacement. Perceptual errors were reduced during movement about a tilted axis on earth. CONCLUSIONS: Reduced perceptual errors during tilts in 1g indicate the importance of otolith and somatosensory cues for maintaining spatial awareness. Tactile cueing may improve navigation in operational environments or clinical populations, providing a non-visual non-auditory feedback of orientation or desired direction heading.

Montgomery, K. L.↗

Perceptual Repetition Blindness Effects

The phenomenon of repetition blindness (RB) may reveal a new limitation on human perceptual processing. Recently, however, researchers have attributed RB to post-perceptual processes such as memory retrieval and/or reporting biases. The standard rapid serial visual presentation (RSVP) paradigm used in most RB studies is, indeed, open to such objections. Here we investigate RB using a "single-frame" paradigm introduced by Johnston and Hale (1984) in which memory demands are minimal. Subjects made only a single judgement about whether one masked target word was the same or different than a post-target probe. Confidence ratings permitted use of signal detection methods to assess sensitivity and bias effects. In the critical condition for RB a precue of the post-target word was provided prior to the target stimulus (identity precue), so that the required judgement amounted to whether the target did or did not repeat the precue word. In control treatments, the precue was either an unrelated word or a dummy.

Hochhaus, Larry↗

Vestibulo-Ocular Responses to Vertical Translation using a Hand-Operated Chair as a Field Measure of Otolith Function

The translational Vestibulo-Ocular Reflex (tVOR) is an important otolith-mediated response to stabilize gaze during natural locomotion. One goal of this study was to develop a measure of the tVOR using a simple hand-operated chair that provided passive vertical motion. Binocular eye movements were recorded with a tight-fitting video mask in ten healthy subjects. Vertical motion was provided by a modified spring-powered chair (swopper.com) at approximately 2 Hz (+/- 2 cm displacement) to approximate the head motion during walking. Linear acceleration was measured with wireless inertial sensors (Xsens) mounted on the head and torso. Eye movements were recorded while subjects viewed near (0.5m) and far (approximately 4m) targets, and then imagined these targets in darkness. Subjects also provided perceptual estimates of target distances. Consistent with the kinematic properties shown in previous studies, the tVOR gain was greater with near targets, and greater with vision than in darkness. We conclude that this portable chair system can provide a field measure of otolith-ocular function at frequencies sufficient to elicit a robust tVOR.

Wood, S. J.↗

Perceptual-components architecture for digital video

A perceptual-components architecture for digital video partitions the image stream into signal components in a manner analogous to that used in the human visual system. These components consist of achromatic and opponent color channels, divided into static and motion channels, further divided into bands of particular spatial frequency and orientation. Bits are allocated to an individual band in accord with visual sensitivity to that band and in accord with the properties of visual masking. This architecture is argued to have desirable features such as efficiency, error tolerance, scalability, device independence, and extensibility.

Watson, Andrew B.↗

A 4800 bps CELP vocoder with an improved excitation

The Stochastic or Code Excited Linear Predictive Coder (CELP) is among the promising candidates for producing good quality speech at low bit rates. However, the speech quality produced suffers from perceived roughness. Many researchers have used pole-zero postfilters to mask the roughness at the output of the synthesis filter. Although the postfilters are effective in masking the noise at low bit rates, they produce spectral distortions. It is proposed that speech can be improved by introducing two modifications to the fixed stochastic codebook. In the first modification, the stochastic codebook is used only when the long term correlations are low. Otherwise, a pulse like codebook is selected. In the second modification, the selected codebook output is weighted using an adaptive spectral shaping procedure. These two modifications were incorporated in a 4800 bps CELP coder and have resulted in a perceptually improved vocoded speech.

Hassanein, Hisham↗

Spatial Standard Observer

The present invention relates to devices and methods for the measurement and/or for the specification of the perceptual intensity of a visual image. or the perceptual distance between a pair of images. Grayscale test and reference images are processed to produce test and reference luminance images. A luminance filter function is convolved with the reference luminance image to produce a local mean luminance reference image . Test and reference contrast images are produced from the local mean luminance reference image and the test and reference luminance images respectively, followed by application of a contrast sensitivity filter. The resulting images are combined according to mathematical prescriptions to produce a Just Noticeable Difference, JND value, indicative of a Spatial Standard Observer. SSO. Some embodiments include masking functions. window functions. special treatment for images lying on or near border and pre-processing of test images.

Watson, Andrw B.↗

Spatial Standard Observer

The present invention relates to devices and methods for the measurement and/or for the specification of the perceptual intensity of a visual image, or the perceptual distance between a pair of images. Grayscale test and reference images are processed to produce test and reference luminance images. A luminance filter function is convolved with the reference luminance image to produce a local mean luminance reference image. Test and reference contrast images are produced from the local mean luminance reference image and the test and reference luminance images respectively, followed by application of a contrast sensitivity filter. The resulting images are combined according to mathematical prescriptions to produce a Just Noticeable Difference, JND value, indicative of a Spatial Standard Observer, SSO. Some embodiments include masking functions, window functions, special treatment for images lying on or near borders and pre-processing of test images.

Watson, Andrew B.↗

Model of visual contrast gain control and pattern masking

We have implemented a model of contrast gain and control in human vision that incorporates a number of key features, including a contrast sensitivity function, multiple oriented bandpass channels, accelerating nonlinearities, and a devisive inhibitory gain control pool. The parameters of this model have been optimized through a fit to the recent data that describe masking of a Gabor function by cosine and Gabor masks [J. M. Foley, "Human luminance pattern mechanisms: masking experiments require a new model," J. Opt. Soc. Am. A 11, 1710 (1994)]. The model achieves a good fit to the data. We also demonstrate how the concept of recruitment may accommodate a variant of this model in which excitatory and inhibitory paths have a common accelerating nonlinearity, but which include multiple channels tuned to different levels of contrast.

NASA Discipline Space Human Factors↗

Vestibular short latency responses to pulsed linear acceleration in unanesthetized animals

Linear acceleration transients were used to elicit vestibular compound action potentials in non-invasively prepared, unanesthetized animals for the first time (chicks, Gallus domesticus, n = 33). Responses were composed of a series of up to 8 dominant peaks occurring within 8 msec of the stimulus. Response amplitudes for 1.0 g stimulus ranged from 1 to 10 microV. A late, slow, triphasic, anesthesia-labile component was identified as a dominant response feature in unanesthetized animals. Amplitudes increased and latencies decreased as stimulus intensity was increased (MANOVA P less than 0.05). Linear regression slope ranges were: amplitudes = 1.0-5.0 microV/g; latencies = -300 to -1100 microseconds/g. Thresholds for single polarity stimuli (0.035 +/- 0.022 g, n = 11) were significantly lower than those of alternating polarity (0.074 +/- 0.028 g, n = 18, P less than 0.001). Bilateral labyrinthectomy eliminated responses whereas bilateral extirpation of cochleae did not significantly change response thresholds. Intense acoustic masking (100/104 dB SL) produced no effect in 2 animals, but did produce small to moderate effects on response amplitudes in 7 others. Changes were attributed to effects on vestibular end organs. Results of unilateral labyrinth blockade (tetrodotoxin) suggest that P1 and N1 preferentially reflect ipsilateral eighth nerve compound action potentials whereas components beyond approximately 2 msec reflect activity from vestibular neurons that depend on both labyrinths. The results demonstrate that short latency vestibular compound action potentials can be measured in unanesthetized, non-invasively prepared animals.

NASA Discipline Neuroscience↗

Evidence Report: Risk of Inadequate Human-Computer Interaction

Human-computer interaction (HCI) encompasses all the methods by which humans and computer-based systems communicate, share information, and accomplish tasks. When HCI is poorly designed, crews have difficulty entering, navigating, accessing, and understanding information. HCI has rarely been studied in an operational spaceflight context, and detailed performance data that would support evaluation of HCI have not been collected; thus, we draw much of our evidence from post-spaceflight crew comments, and from other safety-critical domains like ground-based power plants, and aviation. Additionally, there is a concern that any potential or real issues to date may have been masked by the fact that crews have near constant access to ground controllers, who monitor for errors, correct mistakes, and provide additional information needed to complete tasks. We do not know what types of HCI issues might arise without this "safety net". Exploration missions will test this concern, as crews may be operating autonomously due to communication delays and blackouts. Crew survival will be heavily dependent on available electronic information for just-in-time training, procedure execution, and vehicle or system maintenance; hence, the criticality of the Risk of Inadequate HCI. Future work must focus on identifying the most important contributing risk factors, evaluating their contribution to the overall risk, and developing appropriate mitigations. The Risk of Inadequate HCI includes eight core contributing factors based on the Human Factors Analysis and Classification System (HFACS): (1) Requirements, policies, and design processes, (2) Information resources and support, (3) Allocation of attention, (4) Cognitive overload, (5) Environmentally induced perceptual changes, (6) Misperception and misinterpretation of displayed information, (7) Spatial disorientation, and (8) Displays and controls.

Kritina Holden↗

Full-wave and half-wave rectification in second-order motion perception

Microbalanced stimuli are dynamic displays which do not stimulate motion mechanisms that apply standard (Fourier-energy or autocorrelational) motion analysis directly to the visual signal. In order to extract motion information from microbalanced stimuli, Chubb and Sperling [(1988) Journal of the Optical Society of America, 5, 1986-2006] proposed that the human visual system performs a rectifying transformation on the visual signal prior to standard motion analysis. The current research employs two novel types of microbalanced stimuli: half-wave stimuli preserve motion information following half-wave rectification (with a threshold) but lose motion information following full-wave rectification; full-wave stimuli preserve motion information following full-wave rectification but lose motion information following half-wave rectification. Additionally, Fourier stimuli, ordinary square-wave gratings, were used to stimulate standard motion mechanisms. Psychometric functions (direction discrimination vs stimulus contrast) were obtained for each type of stimulus when presented alone, and when masked by each of the other stimuli (presented as moving masks and also as nonmoving, counterphase-flickering masks). RESULTS: given sufficient contrast, all three types of stimulus convey motion. However, only one-third of the population can perceive the motion of the half-wave stimulus. Observers are able to process the motion information contained in the Fourier stimulus slightly more efficiently than the information in the full-wave stimulus but are much less efficient in processing half-wave motion information. Moving masks are more effective than counterphase masks at hampering direction discrimination, indicating that some of the masking effect is interference between motion mechanisms, and some occurs at earlier stages. When either full-wave and Fourier or half-wave and Fourier gratings are presented simultaneously, there is a wide range of relative contrasts within which the motion directions of both gratings are easily determinable. Conversely, when half-wave and full-wave gratings are combined, the direction of only one of these gratings can be determined with high accuracy. CONCLUSIONS: the results indicate that three motion computations are carried out, any two in parallel: one standard ("first order") and two non-Fourier ("second-order") computations that employ full-wave and half-wave rectification.

Motion Perception/physiology↗

Short latency compound action potentials from mammalian gravity receptor organs

Gravity receptor function was characterized in four mammalian species using far-field vestibular evoked potentials (VsEPs). VsEPs are compound action potentials of the vestibular nerve and central relays that are elicited by linear acceleration ramps applied to the cranium. Rats, mice, guinea pigs, and gerbils were studied. In all species, response onset occurred within 1.5 ms of the stimulus onset. Responses persisted during intense (116 dBSPL) wide-band (50 to 50 inverted question mark omitted inverted question mark000 Hz) forward masking, whereas auditory responses to intense clicks (112 dBpeSPL) were eliminated under the same conditions. VsEPs remained after cochlear extirpation but were eliminated following bilateral labyrinthectomy. Responses included a series of positive and negative peaks that occurred within 8 ms of stimulus onset (range of means at +6 dBre: 1.0 g/ms: P1=908 to 1062 micros, N1=1342 to 1475 micros, P2=1632 to 1952 micros, N2=2038 to 2387 micros). Mean response amplitudes at +6 dBre: 1.0 g/ms ranged from 0.14 to 0.99 microV. VsEP input/output functions revealed latency slopes that varied across peaks and species ranging from -19 to -51 micros/dB. Amplitude-intensity slopes also varied ranging from 0.04 to 0.08 microV/dB for rats and mice. Latency values were comparable to those of birds although amplitudes were substantially smaller in mammals. VsEP threshold values were considerably higher in mammals compared to birds and ranged from -8.1 to -10.5 dBre 1.0 g/ms across species. These results support the hypothesis that mammalian gravity receptors are less sensitive to dynamic stimuli than are those of birds.

Non-NASA Center↗