Search NASASearch

Engineering topics

Rosenholtz, Ruth

Publications and source records attributed to Rosenholtz, Ruth.

A Rorschach Test for Visual Classification Strategies

Contemporary models of pattern, detection and discrimination often employ template matching, but there have been few direct tests of this proposition. Adopting a method developed by Ahumada, we have analyzed how human observers discriminate between two letters of the alphabet ('c' and 'x'). The stimulus consisted of a one degree tall letter plus a four degree field of static white noise, both displayed for 16 frames at a 67 Hz frame rate. Our font and display dimensions approximated those of Solomon and Pelli. The observer identified the letter presented. A QUEST staircase varied letter contrast to maintain a 75% correct rate. For each trial, we preserved the information required to reconstruct the noise field. Possible trial categories based on (signal, response) pairs are: (c,c), (c,x), (x,c), (x,x). Noise fields were averaged separately for each category, and a final classification image was obtained by averaging the four mean images after inverting the sign of categories in which x was the response. If the observer employs a template, it should be revealed in the classification image. The lowpass-filtered classification image derived from 2048 responses of one observer is shown here, along with the corresponding ideal template. An approximation to the ideal template can be seen appropriately located within the classification image. We have also simulated and will discuss the classification images expected from various discrimination models in this experimental context. The construction of classification images appears to be a powerful tool for studying classification strategies used by human observers. Like a Rorschach test, it surreptitiously discovers the inner desires of the visual system.

Watson, Andrew B.

Perceptually-Based Adaptive JPEG Coding

An extension to the JPEG standard (ISO/IEC DIS 10918-3) allows spatial adaptive coding of still images. As with baseline JPEG coding, one quantization matrix applies to an entire image channel, but in addition the user may specify a multiplier for each 8 x 8 block, which multiplies the quantization matrix, yielding the new matrix for the block. MPEG 1 and 2 use much the same scheme, except there the multiplier changes only on macroblock boundaries. We propose a method for perceptual optimization of the set of multipliers. We compute the perceptual error for each block based upon DCT quantization error adjusted according to contrast sensitivity, light adaptation, and contrast masking, and pick the set of multipliers which yield maximally flat perceptual error over the blocks of the image. We investigate the bitrate savings due to this adaptive coding scheme and the relative importance of the different sorts of masking on adaptive coding.

Watson, Andrew B.