Engineering PapersSearch

SEARCH · Engineering Papers

Results for “Multivariate Analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 127 records · Page 7

Structural analysis and design of multivariable control systems: An algebraic approach

The application of algebraic system theory to the design of controllers for multivariable (MV) systems is explored analytically using an approach based on state-space representations and matrix-fraction descriptions. Chapters are devoted to characteristic lambda matrices and canonical descriptions of MIMO systems; spectral analysis, divisors, and spectral factors of nonsingular lambda matrices; feedback control of MV systems; and structural decomposition theories and their application to MV control systems.

Tsay, Yih Tsong

Experiments with a three-dimensional statistical objective analysis scheme using FGGE data

A three-dimensional (3D), multivariate, statistical objective analysis scheme (referred to as optimum interpolation or OI) has been developed for use in numerical weather prediction studies with the FGGE data. Some novel aspects of the present scheme include: (1) a multivariate surface analysis over the oceans, which employs an Ekman balance instead of the usual geostrophic relationship, to model the pressure-wind error cross correlations, and (2) the capability to use an error correlation function which is geographically dependent. A series of 4-day data assimilation experiments are conducted to examine the importance of some of the key features of the OI in terms of their effects on forecast skill, as well as to compare the forecast skill using the OI with that utilizing a successive correction method (SCM) of analysis developed earlier. For the three cases examined, the forecast skill is found to be rather insensitive to varying the error correlation function geographically. However, significant differences are noted between forecasts from a two-dimensional (2D) version of the OI and those from the 3D OI, with the 3D OI forecasts exhibiting better forecast skill. The 3D OI forecasts are also more accurate than those from the SCM initial conditions. The 3D OI with the multivariate oceanic surface analysis was found to produce forecasts which were slightly more accurate, on the average, than a univariate version.

Baker, Wayman E.

A modal analysis of flexible aircraft dynamics with handling qualities implications

A multivariable modal analysis technique is presented for evaluating flexible aircraft dynamics, focusing on meaningful vehicle responses to pilot inputs and atmospheric turbulence. Although modal analysis is the tool, vehicle time response is emphasized, and the analysis is performed on the linear, time-domain vehicle model. In evaluating previously obtained experimental pitch tracking data for a family of vehicle dynamic models, it is shown that flexible aeroelastic effects can significantly affect pitch attitude handling qualities. Consideration of the eigenvalues alone, of both rigid-body and aeroelastic modes, does not explain the simulation results. Modal analysis revealed, however, that although the lowest aeroelastic mode frequency was still three times greater than the short-period frequency, the rigid-body attitude response was dominated by this aeroelastic mode. This dominance was defined in terms of the relative magnitudes of the modal residues in selected vehicle responses.

Schmidt, D. K.

Ensembles of radial basis function networks for spectroscopic detection of cervical precancer

The mortality related to cervical cancer can be substantially reduced through early detection and treatment. However, current detection techniques, such as Pap smear and colposcopy, fail to achieve a concurrently high sensitivity and specificity. In vivo fluorescence spectroscopy is a technique which quickly, noninvasively and quantitatively probes the biochemical and morphological changes that occur in precancerous tissue. A multivariate statistical algorithm was used to extract clinically useful information from tissue spectra acquired from 361 cervical sites from 95 patients at 337-, 380-, and 460-nm excitation wavelengths. The multivariate statistical analysis was also employed to reduce the number of fluorescence excitation-emission wavelength pairs required to discriminate healthy tissue samples from precancerous tissue samples. The use of connectionist methods such as multilayered perceptrons, radial basis function (RBF) networks, and ensembles of such networks was investigated. RBF ensemble algorithms based on fluorescence spectra potentially provide automated and near real-time implementation of precancer detection in the hands of nonexperts. The results are more reliable, direct, and accurate than those achieved by either human experts or multivariate statistical algorithms.

Cervix Neoplasms/pathology

Statistical Evaluation of Time Series Analysis Techniques

The performance of a modified version of NASA's multivariate spectrum analysis program is discussed. A multiple regression model was used to make the revisions. Performance improvements were documented and compared to the standard fast Fourier transform by Monte Carlo techniques.

Benignus, V. A.

Multivariable closed loop control analysis and synthesis for complex flight systems

A flight control system analysis and synthesis method is presented that is intended to be especially suitable for application to vehicles exhibiting complex dynamic characteristics. For such vehicles quantitative handling qualities specifications are not usually available. Howver, handling qualities objectives are specifically introduced in this method via the hypothesis of correlation between pilot ratings and the objective function of an optimal control model of the human pilot. Further, since augmentation and pilot operate in parallel, simultaneous determination of the augmentation and pilot model gains is required. Desirable augmented dynamics are obtained for a variety of complex systems and the method is experimentally verified in the case of simple pilot damper gain selection for optimum pitch tracking performance.

Schmidt, D. K.

A CLIPS expert system for clinical flow cytometry data analysis

An expert system is being developed using CLIPS to assist clinicians in the analysis of multivariate flow cytometry data from cancer patients. Cluster analysis is used to find subpopulations representing various cell types in multiple datasets each consisting of four to five measurements on each of 5000 cells. CLIPS facts are derived from results of the clustering. CLIPS rules are based on the expertise of Drs. Stewart, Duque, and Braylan. The rules incorporate certainty factors based on case histories.

Salzman, G. C.

Generic and ML Workloads in an HPC Datacenter: Node Energy, Job Failures, and Node-Job Analysis

HPC datacenters offer a backbone to the modern digital society. Increasingly, they run Machine Learning (ML) jobs next to generic, compute-intensive workloads, supporting science, business, and other decision-making processes. However, understanding how ML jobs impact the operation of HPC datacenters, relative to generic jobs, remains desirable but understudied. In this work, we leverage long-term operational data, collected from a national-scale production HPC datacenter, and statistically compare how ML and generic jobs can impact the performance, failures, resource utilization, and energy consumption of HPC datacenters. Our study provides key insights, e.g., ML-related power usage causes GPU nodes to run into temperature limitations, median/mean runtime and failure rates are higher for ML jobs than for generic jobs, both ML and generic jobs exhibit highly variable arrival processes and resource demands, significant amounts of energy are spent on unsuccessfully terminating jobs, and concurrent jobs tend to terminate in the same state. We open-source our cleaned-up data traces on Zenodo (https://doi. org/10.5281/zenodo.13685426), and provide our analysis toolkit as software hosted on GitHub (https://github.com/atlarge-research/2024-icpads-hpc-workload-characterization). This study offers multiple benefits for data center administrators, who can improve operational efficiency, and for researchers, who can further improve system designs, scheduling techniques, etc.

crossanalysis

Development and implementation of a low cost micro computer system for LANDSAT analysis and geographic data base applications

Since the implementation of the GRID and IMGRID computer programs for multivariate spatial analysis in the early 1970's, geographic data analysis subsequently moved from large computers to minicomputers and now to microcomputers with radical reduction in the costs associated with planning analyses. Programs designed to process LANDSAT data to be used as one element in a geographic data base were used once NIMGRID (new IMGRID), a raster oriented geographic information system, was implemented on the microcomputer. Programs for training field selection, supervised and unsupervised classification, and image enhancement were added. Enhancements to the color graphics capabilities of the microsystem allow display of three channels of LANDSAT data in color infrared format. The basic microcomputer hardware needed to perform NIMGRID and most LANDSAT analyses is listed as well as the software available for LANDSAT processing.

Faust, N.

Analysis/forecast experiments with a flow-dependent correlation function using FGGE data

The use of a flow-dependent correlation function to improve the accuracy of an optimum interpolation (OI) scheme is examined. The development of the correlation function for the OI analysis scheme used for numerical weather prediction is described. The scheme uses a multivariate surface analysis over the oceans to model the pressure-wind error cross-correlation and it has the ability to use an error correlation function that is flow- and geographically-dependent. A series of four-day data assimilation experiments, conducted from January 5-9, 1979, were used to investigate the effect of the different features of the OI scheme (error correlation) on forecast skill for the barotropic lows and highs. The skill of the OI was compared with that of a successive correlation method (SCM) of analysis. It is observed that the largest difference in the correlation statistics occurred in barotropic and baroclinic lows and highs. The comparison reveals that the OI forecasts were more accurate than the SCM forecasts.

Baker, W. E.

A multiparametric analysis of the Einstein sample of early-type galaxies. 1: Luminosity and ISM parameters

We have conducted bivariate and multivariate statistical analysis of data measuring the luminosity and interstellar medium of the Einstein sample of early-type galaxies (presented by Fabbiano, Kim, & Trinchieri 1992). We find a strong nonlinear correlation between L(sub B) and L(sub X), with a power-law slope of 1.8 +/- 0.1, steepening to 2.0 +/- if we do not consider the Local Group dwarf galaxies M32 and NGC 205. Considering only galaxies with log L(sub X) less than or equal to 40.5, we instead find a slope of 1.0 +/- 0.2 (with or without the Local Group dwarfs). Although E and S0 galaxies have consistent slopes for their L(sub B)-L(sub X) relationships, the mean values of the distribution functions of both L(sub X) and L(sub X)/L(sub B) for the S0 galaxies are lower than those for the E galaxies at the 2.8 sigma and 3.5 sigma levels, respectively. We find clear evidence for a correlation between L(sub X) and the X-ray color C(sub 21), defined by Kim, Fabbiano, & Trinchieri (1992b), which indicates that X-ray luminosity is correlated with the spectral shape below 1 keV in the sense that low-L(sub X) systems have relatively large contributions from a soft component compared with high-L(sub X) systems. We find evidence from our analysis of the 12 micron IRAS data for our sample that our S0 sample has excess 12 micron emission compared with the E sample, scaled by their optical luminosities. This may be due to emission from dust heated in star-forming regions in S0 disks. This interpretation is reinforced by the existence of a strong L(sub 12)-L(sub 100) correlation for our S0 sample that is not found for the E galaxies, and by an analysis of optical-IR colors. We find steep slopes for power-law relationships between radio luminosity and optical, X-ray, and far-IR (FIR) properties. This last point argues that the presence of an FIR-emitting interstellar medium (ISM) in early-type galaxies is coupled to their ability to generate nonthermal radio continuum, as previously argued by, e.g., Walsh et al. (1989). We also find that, for a given L(sub 100), galaxies with larger L(sub X)/L(sub B) tend to be stronger nonthermal radio sources, as originally suggested by Kim & Fabbiano (1990). We note that, while L(sub B) is most strongly correlated with L(sub 6), the total radio luminosity, both L(sub X) and L(sub X)/L(sub B) are more strongly correlated with L(sub 6 CO), the core radio luminosity. These points support the argument (proposed by Fabbiano, Gioia, & Trinchieri 1989) that radio cores in early-type galaxies are fueled by the hot ISM.

Eskridge, Paul B.

Distribution and observed associations of orthostatic blood pressure changes in elderly general medicine outpatients

Factors associated with orthostatic blood pressure change in elderly outpatients were determined by surveying 398 medical clinical outpatients aged 65 years and older. Blood pressure was measured with random-zero sphygmomanometers after patients were 5 minutes in a supine and 5 minutes in a standing position. Orthostatic blood pressure changes were at normally distributed levels with systolic and diastolic pressures dropping an average of 4 mm Hg (standard deviation [SD]=15 mm Hg) and 2 mm Hg (SD=11 mm Hg), respectively. Orthostatic blood pressure changes were unassociated with age, race, sex, body mass, time since eating, symptoms, or other factors. According to multiple linear regression analysis, supine systolic pressure, chronic obstructive pulmonary disease (COPD), and diabetes mellitus were associated with a decrease in systolic pressure on standing. Hypertension, antiarthritic drugs, and abnormal heartbeat were associated with an increase in systolic pressure on standing. For orthostatic diastolic pressure changes, supine diastolic pressure and COPD were associated with a decrease in diastolic pressure on standing. Congestive heart failure was associated with an increase in standing diastolic pressure. Using logistic regression analysis, only supine systolic pressure was associated with a greater than 20-mm Hg drop in systolic pressure (n=53, prevalence=13%). Supine diastolic pressure and COPD were the only variables associated with a greater than 20-mm Hg drop in diastolic pressure (n=16, prevalence=4%). These factors may help physicians in identifying older persons at risk for having orthostatic hypotension.

Non-NASA Center

A method of using cluster analysis to study statistical dependence in multivariate data

A technique is presented that uses both cluster analysis and a Monte Carlo significance test of clusters to discover associations between variables in multidimensional data. The method is applied to an example of a noisy function in three-dimensional space, to a sample from a mixture of three bivariate normal distributions, and to the well-known Fisher's Iris data.

Borucki, W. J.

Regression Model Optimization for the Analysis of Experimental Data

A candidate math model search algorithm was developed at Ames Research Center that determines a recommended math model for the multivariate regression analysis of experimental data. The search algorithm is applicable to classical regression analysis problems as well as wind tunnel strain gage balance calibration analysis applications. The algorithm compares the predictive capability of different regression models using the standard deviation of the PRESS residuals of the responses as a search metric. This search metric is minimized during the search. Singular value decomposition is used during the search to reject math models that lead to a singular solution of the regression analysis problem. Two threshold dependent constraints are also applied. The first constraint rejects math models with insignificant terms. The second constraint rejects math models with near-linear dependencies between terms. The math term hierarchy rule may also be applied as an optional constraint during or after the candidate math model search. The final term selection of the recommended math model depends on the regressor and response values of the data set, the user s function class combination choice, the user s constraint selections, and the result of the search metric minimization. A frequently used regression analysis example from the literature is used to illustrate the application of the search algorithm to experimental data.

Ulbrich, N.

Salvaging Data Records with Missing Data: Data Imputation using the Multivariate t Distribution

When doing multivariate data analysis, one commonobstacle is the presence of incomplete observations, i.e., observationsfor which one or more key fields are blank. Missing datais often countered by deleting entire observations that containmissing data. The negative effects of deleting entire observationsare multiple: deleting observations reduces sample size andcan also result in biased inferences even if data is missing atrandom. In addition, knowledge contained within incompleteobservations is knowledge lost when they are deleted– and theeffort spent collecting that knowledge is effort wasted. Data imputationmethods, or methods of statistically “filling-in” missingdata, can help combat small sample sizes by using the existinginformation in partially complete observations with the end goalof producing less biased and higher confidence inferences. Whena sample from a multivariate normal population is only partiallycomplete, and the missing data meets appropriate assumptions(missing at random), robust data imputation of the missing datacan be implemented with monotone data augmentation (MDA)using the multivariate t distribution.Missing data imputation is applied to data from the NASA InstrumentCost Model (NICM) using the MDA algorithm underthe assumption of having a multivariate t distribution with fixeddegrees of freedom. A sensitivity analysis to the degrees offreedom parameter is presented to demonstrate robustness ofthe multivariate t distribution when dealing with small samplesas compared to the multivariate normal distribution.

DiNicola, Michael

Proceedings of the Third Annual Symposium on Mathematical Pattern Recognition and Image Analysis

Topics addressed include: multivariate spline method; normal mixture analysis applied to remote sensing; image data analysis; classifications in spatially correlated environments; probability density functions; graphical nonparametric methods; subpixel registration analysis; hypothesis integration in image understanding systems; rectification of satellite scanner imagery; spatial variation in remotely sensed images; smooth multidimensional interpolation; and optimal frequency domain textural edge detection filters.

Guseman, L. F., Jr.

A multiparametric analysis of the Einstein sample of early-type galaxies. 2: Galaxy formation history and properties of the interstellar medium

We have conducted bivariate and multivariate statistical analysis of data measuring the integrated luminosity, shape, and potential depth of the Einstein sample of early-type galaxies (presented by Fabbiano et al. 1992). We find significant correlations between the X-ray properties and the axial ratios (a/b) of our sample, such that the roundest systems tend to have the highest L(sub x) and L(sub x)/L(sub B). The most radio-loud objects are also the roundest. We confirm the assertion of Bender et al. (1989) that galaxies with high L(sub x) are boxy (have negative a(sub 4)). Both a/b and a(sub 4) are correlated with L(sub B), but not with IRAS 12 um and 100 um luminosities. There are strong correlations between L(sub x), Mg(sub 2), and sigma(sub nu) in the sense that those systems with the deepest potential wells have the highest L(sub x) and Mg(sub 2). Thus the depth of the potential well appears to govern both the ability to reatin an ISM at the present epoch and to retain the enriched ejecta of early star formation bursts. Both L(sub x)/L(sub B) and L(sub 6) (the 6 cm radio luminosity) show threshold effects with sigma(sub nu) exhibiting sharp increases at log sigma(sub nu) approximately = 2.2. Finally, there is clearly an interrelationship between the various stellar and structural parameters: The scatter in the bivariate relationships between the shape parameters (a/b and a(sub 4)) and the depth parameter sigma(sub nu) is a function of abundance in the sense that, for a given a(sub 4) or a/b, the systems with the highest sigma(sub nu) also have the highest Mg(sub 2). Furthermore, for a constant sigma(sun nu), disky galaxies tend to have higher Mg(sub 2) than boxy ones. Alternatively, for a given abundance, boxy ellipticals tend to be more massive than disky ellipticals. One possibility is that early-type galaxies of a given mass, originating from mergers (boxy ellipticals), have lower abundances than 'primordial' (disky) early-type galaxies. Another is that disky inner isophotes are due not to primordial dissipation collapse, but to either the self-gravitating inner disks of captured spirals or the dissipational collapse of new disk structures from the premerger ISM. The high measured nuclear Mg(sub 2) values would thus be due to enrichment from secondary bursts of star formation triggered by the merging event.

Eskridge, Paul B.