Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “regression analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 127 records · Page 7

Utility of correlation techniques in gravity and magnetic interpretation

Internal correspondence uses Poisson's Theorem in a moving-window linear regression analysis between the anomalous first vertical derivative of gravity and total magnetic field reduced to the pole. The regression parameters provide critical information on source characteristics. The correlation coefficient indicates the strength of the relation between magnetics and gravity. Slope value gives delta j/delta sigma estimates of the anomalous source. The intercept furnishes information on anomaly interference. Cluster analysis consists of the classification of subsets of data into groups of similarity based on correlation of selected characteristics of the anomalies. Model studies are used to illustrate implementation and interpretation procedures of these methods, particularly internal correspondence. Analysis of the results of applying these methods to data from the midcontinent and a transcontinental profile shows they can be useful in identifying crustal provinces, providing information on horizontal and vertical variations of physical properties over province size zones, validating long wavelength anomalies, and isolating geomagnetic field removal problems.

Chandler, V. W.↗

Wildfire Emergency Response Hazard Extraction and Analysis of Trends (HEAT) through Natural Language Processing and Time Series

A methodology for Hazard Extraction and Analysis of Trends (HEAT) is proposed and conducted on a data set of wildfire incident response forms, known as ICS-209-PLUS.The HEAT processes: (1) extract a set of hazards from a data set, (2) calculate hazard-relevant metrics in a primary analysis, (3) analyze trends over time in metrics using timeseries, and (4) examine potential explanations for metric trends using a secondary analysis. Hazards are extracted from narrative data in the ICS-209-PLUS based on a framework previously developed by the authors, using natural language processing. Metrics examined for each hazard include operational time to occurrence, rate of occurrence, frequency, and severity. Primary results include a taxonomy of hazards present in the data set with relevant quantitative metrics. The most frequent hazards identified are environmental and include hazardous terrain. Most hazards occur on average between 35-55% containment. Incidents with hazards tend to have a higher average severity score when compared to the average score for all incidents. Time series of the metrics and relevant predictors, including fire characteristics, fire intensity, and operations, are created to facilitate further analysis. Secondary results used to determine which factors best predict hazard frequency include a correlation matrix and regression analysis. These findings are relevant to safety for current, as well as emerging wildfire operations, and are an exploratory first step in developing historical data-driven risk assessment models.

Sequoia R. Andrade↗

Regression Model Optimization for the Analysis of Experimental Data

A candidate math model search algorithm was developed at Ames Research Center that determines a recommended math model for the multivariate regression analysis of experimental data. The search algorithm is applicable to classical regression analysis problems as well as wind tunnel strain gage balance calibration analysis applications. The algorithm compares the predictive capability of different regression models using the standard deviation of the PRESS residuals of the responses as a search metric. This search metric is minimized during the search. Singular value decomposition is used during the search to reject math models that lead to a singular solution of the regression analysis problem. Two threshold dependent constraints are also applied. The first constraint rejects math models with insignificant terms. The second constraint rejects math models with near-linear dependencies between terms. The math term hierarchy rule may also be applied as an optional constraint during or after the candidate math model search. The final term selection of the recommended math model depends on the regressor and response values of the data set, the user s function class combination choice, the user s constraint selections, and the result of the search metric minimization. A frequently used regression analysis example from the literature is used to illustrate the application of the search algorithm to experimental data.

Ulbrich, N.↗

Theorizing Land Cover and Land Use Change: The Peasant Economy of Colonization in the Amazon Basin

This paper addresses deforestation processes in the Amazon basin. It deploys a methodology combining remote sensing and survey-based fieldwork to examine, with regression analysis, the impact household structure and economic circumstances on deforestation decisions made by colonist farmers in the forest frontiers of Brazil. Unlike most previous regression-based studies, the methodology implemented analyzes behavior at the level of the individual property. The regressions correct for endogenous relationships between key variables, and spatial autocorrelation, as necessary. Variables used in the analysis are specified, in part, by a theoretical development integrating the Chayanovian concept of the peasant household with spatial considerations stemming from von Thuenen. The results from the empirical model indicate that demographic characteristics of households, as well as market factors, affect deforestation in the Amazon. Thus, statistical results from studies that do not include household-scale information may be subject to error. From a policy perspective, the results suggest that environmental policies in the Amazon based on market incentives to small farmers may not be as effective as hoped, given the importance of household factors in catalyzing the demand for land. The paper concludes by noting that household decisions regarding land use and deforestation are not independent of broader social circumstances, and that a full understanding of Amazonian deforestation will require insight into why poor families find it necessary to settle the frontier in the first place.

Caldas, Marcellus↗

Use of ordination and classification procedures to evaluate phytoplankton communities during Superflux II

Cluster analysis and an ordination procedure were performed on two data matrices to investigate real and environmental spatial relationships. Multiple regression analysis was used to relate the measured environmental variables to the phytoplankton community changes. Qualitative type phytoplankton data proved to be less structured in both of these spaces, relative to the biomass data. The salinity gradients of the northern transects covaried significantly with the phytoplankton association changes. In the southern transects the light variable was most important in explaining the variance in the ordination axes. These data suggest the close relationships between phytoplankton community changes and the physical hydrology of the area.

Rutledge, C. K.↗

Computational simulation of probabilistic lifetime strength for aerospace materials subjected to high temperature, mechanical fatigue, creep, and thermal fatigue

The results of a fourth year effort of a research program conducted for NASA-LeRC by The University of Texas at San Antonio (UTSA) are presented. The research included on-going development of methodology that provides probabilistic lifetime strength of aerospace materials via computational simulation. A probabilistic material strength degradation model, in the form of a randomized multifactor interaction equation, is postulated for strength degradation of structural components of aerospace propulsion systems subjected to a number of effects or primitive variables. These primitive variables may include high temperature, fatigue, or creep. In most cases, strength is reduced as a result of the action of a variable. This multifactor interaction strength degradation equation was randomized and is included in the computer program, PROMISC. Also included in the research is the development of methodology to calibrate the above-described constitutive equation using actual experimental materials data together with regression analysis of that data, thereby predicting values for the empirical material constants for each effect or primitive variable. This regression methodology is included in the computer program, PROMISC. Actual experimental materials data were obtained from industry and the open literature for materials typically for applications in aerospace propulsion system components. Material data for Inconel 718 was analyzed using the developed methodology.

Boyce, Lola↗

Computational simulation of probabilistic lifetime strength for aerospace materials subjected to high temperature, mechanical fatigue, creep and thermal fatigue

This report presents the results of a fourth year effort of a research program, conducted for NASA-LeRC by the University of Texas at San Antonio (UTSA). The research included on-going development of methodology that provides probabilistic lifetime strength of aerospace materials via computational simulation. A probabilistic material strength degradation model, in the form of a randomized multifactor interaction equation, is postulated for strength degradation of structural components of aerospace propulsion systems subject to a number of effects or primitive variables. These primitive variables may include high temperature, fatigue or creep. In most cases, strength is reduced as a result of the action of a variable. This multifactor interaction strength degradation equation has been randomized and is included in the computer program, PROMISS. Also included in the research is the development of methodology to calibrate the above-described constitutive equation using actual experimental materials data together with regression analysis of that data, thereby predicting values for the empirical material constants for each effect or primitive variable. This regression methodology is included in the computer program, PROMISC. Actual experimental materials data were obtained from industry and the open literature for materials typically for applications in aerospace propulsion system components. Material data for Inconel 718 has been analyzed using the developed methodology.

Boyce, Lola↗

Changes in aerobic power of men, ages 25-70 yr

This study quantified and compared the cross-sectional and longitudinal influence of age, self-report physical activity (SR-PA), and body composition (%fat) on the decline of maximal aerobic power (VO2peak). The cross-sectional sample consisted of 1,499 healthy men ages 25-70 yr. The 156 men of the longitudinal sample were from the same population and examined twice, the mean time between tests was 4.1 (+/- 1.2) yr. Peak oxygen uptake was determined by indirect calorimetry during a maximal treadmill exercise test. The zero-order correlations between VO2peak and %fat (r = -0.62) and SR-PA (r = 0.58) were significantly (P < 0.05) higher that the age correlation (r = -0.45). Linear regression defined the cross-sectional age-related decline in VO2peak at 0.46 ml.kg-1.min-1.yr-1. Multiple regression analysis (R = 0.79) showed that nearly 50% of this cross-sectional decline was due to %fat and SR-PA, adding these lifestyle variables to the multiple regression model reduced the age regression weight to -0.26 ml.kg-1.min-1.yr-1. Statistically controlling for time differences between tests, general linear models analysis showed that longitudinal changes in aerobic power were due to independent changes in %fat and SR-PA, confirming the cross-sectional results.

Oxygen Consumption/physiology↗

Comparison of regression modeling techniques for resource estimation

The development and validation of resource utilization models is an active area of software engineering research. Regression analysis is the principal tool employed in these studies. However, little attention was given to determining which of the various regression methods available is the most appropriate. The objective of the study is to compare three alternative regession procedures by examining the results of their application to one commonly accepted equations for resource estimation. The data studied was summarized, the resource estimation equation was described, the regression procedures were explained, and the results obtained from the proceures were compared.

Card, D. N.↗

Assessing the Transferability of Statistical Predictive Models for Leaf Area Index Between Two Airborne Discrete Return LiDAR Sensor Designs Within Multiple Intensely Managed Loblolly Pine Forest Locations in the South-Eastern USA

Leaf area is an important forest structural variable which serves as the primary means of mass and energy exchange within vegetated ecosystems. The objective of the current study was to determine if leaf area index (LAI) could be estimated accurately and consistently in five intensively managed pine plantation forests using two multiple-return airborne LiDAR datasets. Field measurements of LAI were made using the LiCOR LAI2000 and LAI2200 instruments within 116 plots were established of varying size and within a variety of stand conditions (i.e. stand age, nutrient regime and stem density) in North Carolina and Virginia in 2008 and 2013. A number of common LiDAR return height and intensity distribution metrics were calculated (e.g. average return height), in addition to ten indices, with two additional variants, utilized in the surrounding literature which have been used to estimate LAI and fractional cover, were calculated from return heights and intensity, for each plot extent. Each of the indices was assessed for correlation with each other, and was used as independent variables in linear regression analysis with field LAI as the dependent variable. All LiDAR derived metrics were also entered into a forward stepwise linear regression. The results from each of the indices varied from an R2 of 0.33 (S.E. 0.87) to 0.89 (S.E. 0.36). Those indices calculated using ratios of all returns produced the strongest correlations, such as the Above and Below Ratio Index (ABRI) and Laser Penetration Index 1 (LPI1). The regression model produced from a combination of three metrics did not improve correlations greatly (R2 0.90; S.E. 0.35). The results indicate that LAI can be predicted over a range of intensively managed pine plantation forest environments accurately when using different LiDAR sensor designs. Those indices which incorporated counts of specific return numbers (e.g. first returns) or return intensity correlated poorly with field measurements. There were disparities between the number of different types of returns and intensity values when comparing the results from two LiDAR sensors, indicating that predictive models developed using such metrics are not transferable between datasets with different acquisition parameters. Each of the indices were significantly correlated with one another, with one exception (LAI proxy), in particular those indices calculated from all returns, which indicates similarities in information content for those indices. It can then be argued that LiDAR indices have reached a similar stage in development to those calculated from optical-spectral sensors, but which offer a number of advantages, such as the reduction or removal of saturation issues in areas of high biomass.

Sumnall, Matthew↗

Molecular Static Third-Order Polarizabilities of Carbon-Cage Fullerene and Their Correlation with Three Geometric Properties: Symmetry, Aromaticity, and Size

The static third-order polarizabilities (gamma) of C60, C70, five isomers of C78 and two isomers of C84 were analyzed in terms of three properties, from a geometric point of view: symmetry, aromaticity and size. The polarizability values were based on the finite field approximation using a semiempirical Hamiltonian (AM1) and applied to molecular structures obtained from density functional theory calculations. Symmetry was characterized by the molecular group order. The selection of 6-member rings as aromatic was determined from an analysis of bond lengths. Maximum interatomic distance and surface area were the parameters considered with respect to size. Based on triple linear regression analysis, it was found that the static linear polarizability (alpha) and gamma in these molecules respond differently to geometrical properties: alpha depends almost exclusively on surface area while gamma is affected by a combination of number of aromatic rings, length and group order, in decreasing importance. In the case of alpha, valence electron contributions provide the same information as all-electron estimates. For gamma, the best correlation coefficients are obtained when all-electron estimates are used and when the dependent parameter is ln(gamma) instead of gamma.

Moore, C. E.↗

Using Knowledge Analytics to Search and Characterize Mass Properties Aerospace Data

There is growing capability in the field of “Big Data” and “Data Analytics” which Mass Properties Engineers might like to take advantage of. This paper utilizes an implementation of the IBM Knowledge Analytics and Watson search capabilities to explore a corpus of material developed primarily with the interests of Mass Properties Engineers and vehicle concept developers at its forefront. The full collection of SAWE (Society of Allied Weight Engineers, Inc.) Technical Papers from 1939 through 2015 is a major portion of the knowledge content. Additional aerospace vehicle design information includes metadata from AIAA (American Institute for Aeronautics and Astronautics), and INCOSE (International Council on Systems Engineering) as well as author-provided personal search material. This data is processed with respect to certain expected content, data taxonomies and key words to become the core data in NASA Langley Research Center’s “Vehicle Analysis Analytics”, IBM Watson Content. Processed data becomes the corpus of information which is interrogated to provide examples of finding data for mass regression analysis, technology impacts on MPE (Mass Properties Engineering), mass properties control, standards, and other aspects of interest.

Cerro, Jeffrey A.↗

An Airline-Based Multilevel Analysis of Airfare Elasticity for Passenger Demand

Price elasticity of passenger demand for a specific airline is estimated. The main drivers affecting passenger demand for air transportation are identified. First, an Ordinary Least Squares regression analysis is performed. Then, a multilevel analysis-based methodology to investigate the pattern of variation of price elasticity of demand among the various routes of the airline under study is proposed. The experienced daily passenger demands on each fare-class are grouped for each considered route. 9 routes were studied for the months of February and May in years from 1999 to 2002, and two fare-classes were defined (business and economy). The analysis has revealed that the airfare elasticity of passenger demand significantly varies among the different routes of the airline.

Castelli, Lorenzo↗

A Comparison of Two Balance Calibration Model Building Methods

Simulated strain-gage balance calibration data is used to compare the accuracy of two balance calibration model building methods for different noise environments and calibration experiment designs. The first building method obtains a math model for the analysis of balance calibration data after applying a candidate math model search algorithm to the calibration data set. The second building method uses stepwise regression analysis in order to construct a model for the analysis. Four balance calibration data sets were simulated in order to compare the accuracy of the two math model building methods. The simulated data sets were prepared using the traditional One Factor At a Time (OFAT) technique and the Modern Design of Experiments (MDOE) approach. Random and systematic errors were introduced in the simulated calibration data sets in order to study their influence on the math model building methods. Residuals of the fitted calibration responses and other statistical metrics were compared in order to evaluate the calibration models developed with different combinations of noise environment, experiment design, and model building method. Overall, predicted math models and residuals of both math model building methods show very good agreement. Significant differences in model quality were attributable to noise environment, experiment design, and their interaction. Generally, the addition of systematic error significantly degraded the quality of calibration models developed from OFAT data by either method, but MDOE experiment designs were more robust with respect to the introduction of a systematic component of the unexplained variance.

DeLoach, Richard↗

A data–model approach to interpreting speleothem oxygen isotope records from monsoon regions

Reconstruction of past changes in monsoon climate from speleothem oxygen isotope (δ 18 O) records is complex because δ 18 O signals can be influenced by multiple factors including changes in precipitation, precipitation recycling over land, temperature at the moisture source, and changes in the moisture source region and transport pathway. Here, we analyse >150 speleothem records of the Speleothem Isotopes Synthesis and AnaLysis (SISAL) database to produce composite regional trends in δ 18 O in monsoon regions; compositing minimises the influence of site-specific karst and cave processes that can influence individual site records. We compare speleothem δ 18 O observations with isotope-enabled climate model simulations to investigate the specific climatic factors causing these regional trends. We focus on differences in δ 18 O signals between the mid-Holocene, the peak of the Last Interglacial (Marine Isotope Stage 5e) and the Last Glacial Maximum as well as on δ 18 O evolution through the Holocene. Differences in speleothem δ 18 O between the mid-Holocene and the Last Interglacial in the East Asian and Indian monsoons are small, despite the larger summer insolation values during the Last Interglacial. Last Glacial Maximum δ 18 O values are significantly less negative than interglacial values. Comparison with simulated glacial–interglacial δ 18 O shows that changes are principally driven by global shifts in temperature and regional precipitation. Holocene speleothem δ 18 O records show distinct and coherent regional trends. Trends are similar to summer insolation in India, China and southwestern South America, but they are different in the Indonesian–Australian region. Redundancy analysis shows that 37 % of Holocene variability can be accounted for by latitude and longitude, supporting the differentiation of records into individual monsoon regions. Regression analysis of simulated precipitation δ 18 O and climate variables show significant relationships between global Holocene monsoon δ 18 O trends and changes in precipitation, atmospheric circulation and (to a lesser extent) source area temperature, whereas precipitation recycling is non-significant. However, there are differences in regional-scale mechanisms: there are clear relationships between changes in precipitation and δ 18 O for India, southwestern South America and the Indonesian–Australian regions but not for the East Asian monsoon. Changes in atmospheric circulation contribute to δ 18 O trends in the East Asian, Indian and Indonesian–Australian monsoons, and a weak source area temperature effect is observed over southern and central America and Asia. Precipitation recycling is influential in southwestern South America and southern Africa. Overall, our analyses show that it is possible to differentiate the impacts of specific climatic mechanisms influencing precipitation δ 18 O and use this analysis to interpret changes in speleothem δ 18 O.

speleothem oxygen isotope records↗

Wind Tunnel Strain-Gage Balance Calibration Data Analysis Using a Weighted Least Squares Approach

A new approach is presented that uses a weighted least squares fit to analyze wind tunnel strain-gage balance calibration data. The weighted least squares fit is specifically designed to increase the influence of single-component loadings during the regression analysis. The weighted least squares fit also reduces the impact of calibration load schedule asymmetries on the predicted primary sensitivities of the balance gages. A weighting factor between zero and one is assigned to each calibration data point that depends on a simple count of its intentionally loaded load components or gages. The greater the number of a data point's intentionally loaded load components or gages is, the smaller its weighting factor becomes. The proposed approach is applicable to both the Iterative and Non-Iterative Methods that are used for the analysis of strain-gage balance calibration data in the aerospace testing community. The Iterative Method uses a reasonable estimate of the tare corrected load set as input for the determination of the weighting factors. The Non-Iterative Method, on the other hand, uses gage output differences relative to the natural zeros as input for the determination of the weighting factors. Machine calibration data of a six-component force balance is used to illustrate benefits of the proposed weighted least squares fit. In addition, a detailed derivation of the PRESS residuals associated with a weighted least squares fit is given in the appendices of the paper as this information could not be found in the literature. These PRESS residuals may be needed to evaluate the predictive capabilities of the final regression models that result from a weighted least squares fit of the balance calibration data.

calibration analysis↗

Multivariate Analysis, Retrieval, and Storage System (MARS). Volume 6: MARS System - A Sample Problem (Gross Weight of Subsonic Transports)

The Mars system is a tool for rapid prediction of aircraft or engine characteristics based on correlation-regression analysis of past designs stored in the data bases. An example of output obtained from the MARS system, which involves derivation of an expression for gross weight of subsonic transport aircraft in terms of nine independent variables is given. The need is illustrated for careful selection of correlation variables and for continual review of the resulting estimation equations. For Vol. 1, see N76-10089.

Hague, D. S.↗

General purpose research rotor

An analytical study, under a NASA contract, is performed on an advanced flight research rotor (four-bladed, 54 ft in diameter, with bearingless rotor retention characteristics) to determine the sensitivity of total rotor characteristics such as vibratory hub loads, rotor horsepower, and blade loads, to parametric variability of the rotor configuration. The sensitivity of the rotor to various combinations of blade planform taper, percent of blade span that is tapered, tip sweep angle, built-in-twist, and torsional frequency is determined for specific configurations by randomly selecting combinations of these parameters. Characteristics of other intermediate rotor configurations are determined by a regression analysis. The results show that a wide range of rotor total performance characteristics can be obtained for a rotor trimmed to the same flight conditions. The regression equations predict total performance of the rotor very well and appear to be a useful analytical tool for rotor design optimization. Figures showing the results of the various tests are given along with a table of the regression coefficients.

Jones, R.↗