Engineering PapersSearch

Engineering topics

Null, Cynthia H.

Publications and source records attributed to Null, Cynthia H..

At least 55 records · Page 3

Downsizing Visualization Platforms: From Crays to Indigos and Beyond

Despite its increasingly negative connotation in other domains, "downsizing" is a positive trend as it applies to the computer hardware necessary to perform scientific visualization. In this talk, we will consider how the visualization of a particular data set, the Digital Terrain Model (DTM) derived from the Viking Orbiter imagery, has been realized in three distinct projects over the past decade. These examples serve to demonstrate how the vast improvements in computational performance both decrease the cost of such visualization efforts, and permit an increasing level of interactivity. We then consider how even today's graphical systems require the visualization designer to make intelligent choices and trade-offs in database rendering. Finally, we discuss how insights from an understanding of human visual perception can guide these design decisions, and suggest new options for visualization hardware and software.

Kaiser, Mary K.

Smooth Pursuit of a Partially Occluded Object

There has long been qualitative evidence that humans can pursue an object defined only by the motion of its parts. We explored this quantitatively using an occluded diamond stimulus. Four subjects (one naive) tracked a line-figure diamond moving along an elliptical path (0.9 Hz) either clockwise (CW) or counterclockwise (CCW) behind either an X-shaped aperture (CROSS) or two vertical rectangular apertures (BARS), which obscured the corners. Although the stimulus consisted of only four line segments (108 cd/square m) moving within a visible aperture (0.2 cd/square m) behind a foreground (38 cd/square m), it is largely perceived as a coherently moving diamond. The inter-saccadic portions of eye-position traces were fit with sinusoids. All subjects tracked object motion with considerable temporal accuracy. The mean phase lag was 5 deg/6 deg (CROSS/BARS) and the mean relative phase between the horizontal and vertical components was +95 deg/+92 deg (CW) and -85 deg/-75 deg (CCW), which is close to perfect. Furthermore, a chi-square analysis showed that 56% of BARS trials were consistent with tracking the correct elliptical shape (p is less than 0.05), although segment motion was purely vertical. These data disprove the main tenet of most models of pursuit: that it is a system that seeks to minimize retinal image motion through negative feedback. Rather, the main drive must be a visual signal which has already integrated spatiotemporal retinal information into an object-motion signal.

Stone, L. S.

Design of Multi-Parameter Steerable Functions Using Cascade Basis Reduction

A new cascade basis reduction method of computing the optimal least-squares set of basis functions steering a given function is presented. The method combines the Lie group-theoretic and the singular value decomposition approaches in such a way that their respective strengths complement each other. Since the Lie group-theoretic approach is used, the set of basis and steering functions computed can be expressed analytically. Because the singular value decomposition method is used, this set of basis and steering functions is optimal in the least-squares sense. Furthermore, the computational complexity in designing basis functions for transformation groups with large numbers of parameters is significantly reduced. The efficiency of the cascade basis reduction method is demonstrated by designing a set of basis functions that steers a Gabor function under the four-parameter linear transformation group.

Teo, P.

Translation and Rotation Trade Off in Human Visual Heading Estimation

We have previously shown that, during simulated curvilinear motion, humans can make reasonably accurate and precise heading judgments from optic flow without either oculomotor or static-depth cues about rotation. We now systematically investigate the effect of varying the parameters of self-motion. We visually simulated 400 ms of self-motion along curved paths (constant rotation and translation rates, fixed retinocentric heading) towards two planes of random dots at 10.3 m and 22.3 m at mid-trial. Retinocentric heading judgments of 4 observers (2 naive) were measured for 12 different combinations of translation (T between 4 and 16 m/s) and rotation (R either 8 or 16 deg/s). In the range tested, heading bias and uncertainty decrease quasilinearly with T/R, but the bias also appears to depend on R. If depth is held constant, the ratio T/R can account for much of the variation in the accuracy and precision of human visual heading estimation, although further experiments are needed to resolve whether absolute rotation rate, total flow rate, or some other factor can account for the observed -2 deg shift between the bias curves.

Stone, Leland S.

Perceptual Classification Images from Vernier Acuity Masked by Noise

Letting external noise rather than internal noise limit discrimination performance allows information to be extracted about the observer's stimulus classification rule. A perceptual classification image is the correlation over trials between the noise amplitude at a spatial location and the observer's responses. If, for example, the observer followed the rule of the ideal observer, the perceptual classification image would be an estimate of the ideal observer filter, the difference between the two unmasked images being discriminated. Perceptual classification images were estimated for a vernier discrimination task. The display screen had 48 pixels per degree horizontally and vertically. The no-offset image had a dark horizontal line of 4 pixels, a 1 pixel space, and 4 more dark pixels. Classification images were based on 1600 discrimination trials with the line contrast adjusted to keep the error rate near 25 percent. In the offset image, the second line was one pixel higher. Unlike the ideal observer filter (a horizontal dipole), the observer perceptual classification images are strongly oriented. Fourier transforms of the classification images had a peak amplitude near one cycle per degree and an orientation near 25 degrees. The spatial spread is much more than image blur predicts, and probably indicates the spatial position uncertainty in the task.

Ahumada, A. J.

A Rorschach Test for Visual Classification Strategies

Contemporary models of pattern, detection and discrimination often employ template matching, but there have been few direct tests of this proposition. Adopting a method developed by Ahumada, we have analyzed how human observers discriminate between two letters of the alphabet ('c' and 'x'). The stimulus consisted of a one degree tall letter plus a four degree field of static white noise, both displayed for 16 frames at a 67 Hz frame rate. Our font and display dimensions approximated those of Solomon and Pelli. The observer identified the letter presented. A QUEST staircase varied letter contrast to maintain a 75% correct rate. For each trial, we preserved the information required to reconstruct the noise field. Possible trial categories based on (signal, response) pairs are: (c,c), (c,x), (x,c), (x,x). Noise fields were averaged separately for each category, and a final classification image was obtained by averaging the four mean images after inverting the sign of categories in which x was the response. If the observer employs a template, it should be revealed in the classification image. The lowpass-filtered classification image derived from 2048 responses of one observer is shown here, along with the corresponding ideal template. An approximation to the ideal template can be seen appropriately located within the classification image. We have also simulated and will discuss the classification images expected from various discrimination models in this experimental context. The construction of classification images appears to be a powerful tool for studying classification strategies used by human observers. Like a Rorschach test, it surreptitiously discovers the inner desires of the visual system.

Watson, Andrew B.

DCTune Perceptual Optimization of Compressed Dental X-Rays

In current dental practice, x-rays of completed dental work are often sent to the insurer for verification. It is faster and cheaper to transmit instead digital scans of the x-rays. Further economies result if the images are sent in compressed form. DCTune is a technology for optimizing DCT (digital communication technology) quantization matrices to yield maximum perceptual quality for a given bit-rate, or minimum bit-rate for a given perceptual quality. Perceptual optimization of DCT color quantization matrices. In addition, the technology provides a means of setting the perceptual quality of compressed imagery in a systematic way. The purpose of this research was, with respect to dental x-rays, 1) to verify the advantage of DCTune over standard JPEG (Joint Photographic Experts Group), 2) to verify the quality control feature of DCTune, and 3) to discover regularities in the optimized matrices of a set of images. We optimized matrices for a total of 20 images at two resolutions (150 and 300 dpi) and four bit-rates (0.25, 0.5, 0.75, 1.0 bits/pixel), and examined structural regularities in the resulting matrices. We also conducted psychophysical studies (1) to discover the DCTune quality level at which the images became 'visually lossless,' and (2) to rate the relative quality of DCTune and standard JPEG images at various bitrates. Results include: (1) At both resolutions, DCTune quality is a linear function of bit-rate. (2) DCTune quantization matrices for all images at all bitrates and resolutions are modeled well by an inverse Gaussian, with parameters of amplitude and width. (3) As bit-rate is varied, optimal values of both amplitude and width covary in an approximately linear fashion. (4) Both amplitude and width vary in systematic and orderly fashion with either bit-rate or DCTune quality; simple mathematical functions serve to describe these relationships. (5) In going from 150 to 300 dpi, amplitude parameters are substantially lower and widths larger at corresponding bit-rates or qualities. (6) Visually lossless compression occurs at a DCTune quality value of about 1. (7) At 0.25 bits/pixel, comparative ratings give DCTune a substantial advantage over standard JPEG. As visually lossless bit-rates are approached, this advantage of necessity diminishes. We have concluded that DCTune optimized quantization matrices provide better visual quality than standard JPEG. Meaningful quality levels may be specified by means of the DCTune metric. Optimized matrices are very similar across the class of dental x-rays, suggesting the possibility of a 'class-optimal' matrix. DCTune technology appears to provide some value in the context of compressed dental x-rays.

Watson, Andrew B.

Eye Movements Reveal Hierarchical Motion Processing

Purpose: In the analysis of visual motion, local features such as orientation are analyzed early in the cortical processing stream (V1), while integration across orientation and space is thought to occur in higher cortical areas such as MT, MST, etc. If all areas provide inputs to eye movement control centers, we would expect that local properties would drive eye movements with relatively short latencies, while global properties would require longer latencies. When such latencies are observed, they can provide information about when (and where?) various stimulus properties are analyzed. Methods: The stimulus employed was an elliptical Gabor patch with a drifting carrier, in which the orientations of the carrier grating and the contrast window were varied independently. We have previously demonstrated that the directional percepts evoked by this stimulus vary between the "grating direction" (the normal to the grating's orientation) and the "window direction", and that similar effects can be observed in reflexive eye movements. Subjects viewed such a stimulus while attempting to maintain steady fixation on the center of the pattern, and the small reflexive eye movements ("stare OKN") were recorded. In the middle of the trial, the orientation of either the grating or the window was rotated smoothly by 30 degrees. Results: Responses to the shift of both grating orientation and window orientation are seen in the average OKN slow phase velocity. Grating rotations produce a rapid OKN rotation to the grating direction (100 ms latency, 300 ms time constant), followed by a slower rebound to the steady state perceived direction midway between the grating and window directions. Window rotations, on the other hand, evoke a slower response (200 ms latency, 500 ms time constant). Conclusions: The results demonstrate multiple cortical inputs to eye movement control: a fast, early input driven by orientation, and a slower input from higher areas sensitive to global stimulus properties.

Mulligan, Jeffrey B.

Virtual Acoustics, Aeronautics and Communications

An optimal approach to auditory display design for commercial aircraft would utilize both spatialized ("3-D") audio techniques and active noise cancellation for safer operations. Results from several aircraft simulator studies conducted at NASA Ames Research Center are reviewed, including Traffic alert and Collision Avoidance System (TCAS) warnings, spoken orientation "beacons" for gate identification and collision avoidance on the ground, and hardware for improved speech intelligibility. The implications of hearing loss amongst pilots is also considered.

Begault, Durand R.

The Role of Response Processing in the Single-Channel PRP Bottleneck

Does the single-channel bottleneck occur in the psychological refractory period paradigm in the absence of overt responses? A go/no-go paradigm was used to vary a default response on all trials, multiple go responses were used. Results show that the bottleneck is still present on trials that do not require a response, but its duration is reduced.

VanSelst, Mark

Perceptual Image Compression in Telemedicine

The next era of space exploration, especially the "Mission to Planet Earth" will generate immense quantities of image data. For example, the Earth Observing System (EOS) is expected to generate in excess of one terabyte/day. NASA confronts a major technical challenge in managing this great flow of imagery: in collection, pre-processing, transmission to earth, archiving, and distribution to scientists at remote locations. Expected requirements in most of these areas clearly exceed current technology. Part of the solution to this problem lies in efficient image compression techniques. For much of this imagery, the ultimate consumer is the human eye. In this case image compression should be designed to match the visual capacities of the human observer. We have developed three techniques for optimizing image compression for the human viewer. The first consists of a formula, developed jointly with IBM and based on psychophysical measurements, that computes a DCT quantization matrix for any specified combination of viewing distance, display resolution, and display brightness. This DCT quantization matrix is used in most recent standards for digital image compression (JPEG, MPEG, CCITT H.261). The second technique optimizes the DCT quantization matrix for each individual image, based on the contents of the image. This is accomplished by means of a model of visual sensitivity to compression artifacts. The third technique extends the first two techniques to the realm of wavelet compression. Together these two techniques will allow systematic perceptual optimization of image compression in NASA imaging systems. Many of the image management challenges faced by NASA are mirrored in the field of telemedicine. Here too there are severe demands for transmission and archiving of large image databases, and the imagery is ultimately used primarily by human observers, such as radiologists. In this presentation I will describe some of our preliminary explorations of the applications of our technology to the special problems of telemedicine.

Watson, Andrew B.

Moving Through Time: The Utility of a Temporal Metric for Vehicular Control

Recent work on time-to-contact (TTC) and time-to-passage (TTP) estimation have revealed conflicting evidence. On the one hand, when tau information is available, observers often fail to use it properly. For instance, absolute size of the object, rotation, and contrast all interfere with estimation accuracy. That is, the visual system often is unable to isolate tau. On the other hand, judgments are remarkably robust when invariant information is no longer available. Observers seem to adjust to many disturbances. This symposium intends to stimulate a discussion of recent findings within visual and auditory TTC paradigms. Alternative concepts that are able to accommodate these findings should be entertained and, if found valid, incorporated into the classic theory of temporal-range estimation.

Kaiser, Mary K.

Parafoveal Target Detectability Reversal Predicted by Local Luminance and Contrast Gain Control

This project is part of a program to develop image discrimination models for the prediction of the detectability of objects in a range of backgrounds. We wanted to see if the models could predict parafoveal object detection as well as they predict detection in foveal vision. We also wanted to make our simplified models more general by local computation of luminance and contrast gain control. A signal image (0.78 x 0.17 deg) was made by subtracting a simulated airport runway scene background image (2.7 deg square) from the same scene containing an obstructing aircraft. Signal visibility contrast thresholds were measured in a fully crossed factorial design with three factors: eccentricity (0 deg or 4 deg), background (uniform or runway scene background), and fixed-pattern white noise contrast (0%, 5%, or 10%). Three experienced observers responded to three repetitions of 60 2IFC trials in each condition and thresholds were estimated by maximum likelihood probit analysis. In the fovea the average detection contrast threshold was 4 dB lower for the runway background than for the uniform background, but in the parafovea, the average threshold was 6 dB higher for the runway background than for the uniform background. This interaction was similar across the different noise levels and for all three observers. A likely reason for the runway background giving a lower threshold in the fovea is the low luminance near the signal in that scene. In our model, the local luminance computation is controlled by a spatial spread parameter. When this parameter and a corresponding parameter for the spatial spread of contrast gain were increased for the parafoveal predictions, the model predicts the interaction of background with eccentricity.

Ahumada, Albert J., Jr.

Tracking Virtual Trajectories

Current models of smooth pursuit eye movements assume that it is largely driven by retinal image motion. We tested this hypothesis by measuring pursuit of elliptical motion (3.2s, 0.9 Hz, 1.4 deg x 1.6 deg, 4 randomly interleaved phases) of either a small spot ("real" motion) or of a line-figure diamond viewed through apertures such that only the motion of four isolated oblique line segments was visible ("virtual" motion). Each segment moved sinusoidally along a linear trajectory yet subjects perceived a diamond moving along an elliptical path behind the aperture. We found, as expected, that real motion produced accurate tracking (N = 2) with mean gain (over horizontal and vertical) of 0.9, mean phase of -6 deg (lag), mean relative phase (H vs V) of 90 +/- 8 deg (RMS error). Virtual motion behind an X-shaped aperture (N= 4 with one naive) yielded a mean gain of 0.7, mean phase of -11 deg, mean relative phase of 87 +/- 15 deg. We also measured pursuit with the X-shaped aperture using a higher segment luminance which prevents the segments from being grouped into a coherently moving diamond while keeping the motion otherwise identical. In this incoherent case, the same four subjects no longer showed consistent elliptical tracking (RMS error in relative phase rose to 60 deg) suggesting that perceptual coherence is critical. Furthermore, to rule out tracking of the centroid, we also used vertical apertures so that all segment motion was vertical (N = 3). This stimulus still produced elliptical tracking (mean relative phase of 84 +/- 19 deg), albeit with a lower gain (0.6). These data show that humans can track moving objects reasonably accurately even when the trajectory can only be derived by spatial integration of motion signals. Models that merely seek to minimize retinal or local stimulus motion cannot explain these results.

Stone, Leland S.

Expectancy and Repetition in Task Preparation

We studied the mechanisms of task preparation using a design that pitted task expectancy against task repetition. In one experiment, two simple cognitive tasks were presented in a predictable sequence containing both repetitions and non-repetitions. The typical task sequence was AABBAABB. Occasional violations of this sequence allowed us to measure the effects of valid versus invalid expectancy. With this design, we were able to study the effects of task expectancy, task repetition, and interaction.

Ruthruff, E.

Does Involuntary Attentional Capture Curtail Stimulus Processing

Under certain conditions uninformative visual cues at non-target locations increase target RTs. The typical explanation is that visual attention has been involuntarily summoned to the cued location and is unavailable when needed at the target location. Here we present evidence that capture has not just oriented attention away from the target location, but has lead to the processing of stimuli at the location of the uninformative cue.

Remington, Roger W.

Perceptually-Based Adaptive JPEG Coding

An extension to the JPEG standard (ISO/IEC DIS 10918-3) allows spatial adaptive coding of still images. As with baseline JPEG coding, one quantization matrix applies to an entire image channel, but in addition the user may specify a multiplier for each 8 x 8 block, which multiplies the quantization matrix, yielding the new matrix for the block. MPEG 1 and 2 use much the same scheme, except there the multiplier changes only on macroblock boundaries. We propose a method for perceptual optimization of the set of multipliers. We compute the perceptual error for each block based upon DCT quantization error adjusted according to contrast sensitivity, light adaptation, and contrast masking, and pick the set of multipliers which yield maximally flat perceptual error over the blocks of the image. We investigate the bitrate savings due to this adaptive coding scheme and the relative importance of the different sorts of masking on adaptive coding.

Watson, Andrew B.

A Virtual Audio Guidance and Alert System for Commercial Aircraft Operations

Our work in virtual reality systems at NASA Ames Research Center includes the area of aurally-guided visual search, using specially-designed audio cues and spatial audio processing (also known as virtual or "3-D audio") techniques (Begault, 1994). Previous studies at Ames had revealed that use of 3-D audio for Traffic Collision Avoidance System (TCAS) advisories significantly reduced head-down time, compared to a head-down map display (0.5 sec advantage) or no display at all (2.2 sec advantage) (Begault, 1993, 1995; Begault & Pittman, 1994; see Wenzel, 1994, for an audio demo). Since the crew must keep their head up and looking out the window as much as possible when taxiing under low-visibility conditions, and the potential for "blunder" is increased under such conditions, it was sensible to evaluate the audio spatial cueing for a prototype audio ground collision avoidance warning (GCAW) system, and a 3-D audio guidance system. Results were favorable for GCAW, but not for the audio guidance system.

Begault, Durand R.