Engineering PapersSearch

SEARCH · Engineering Papers

Results for “hypothesis testing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Multiple object tracking with non-unique data-to-object association via generalized hypothesis testing

A generalized hypothesis testing approach is applied to the problem of tracking several objects where several different associations of data with objects are possible. Such problems occur, for instance, when attempting to distinctly track several aircraft maneuvering near each other or when tracking ships at sea. Conceptually, the problem is solved by first, associating data with objects in a statistically reasonable fashion and then, tracking with a bank of Kalman filters. The objects are assumed to have motion characterized by a fixed but unknown deterministic portion plus a random process portion modeled by a shaping filter. For example, the object might be assumed to have a mean straight line path about which it maneuvers in a random manner. Several hypothesized associations of data with objects are possible because of ambiguity as to which object the data comes from, false alarm/detection errors, and possible uncertainty in the number of objects being tracked. The statistical likelihood function is computed for each possible hypothesized association of data with objects. Then the generalized likelihood is computed by maximizing the likelihood over parameters that define the deterministic motion of the object.

Porter, D. W.

Pilot model hypothesis testing

The aircraft control time history predicted by the optimal control pilot model and actual pilot tracking data obtained from NASA Langley's differential maneuvering simulator (DMS) are analyzed. The analysis is performed using a hypothesis testing scheme modified to allow for changes in the true hypothesis. A finite number of pilot models, each with different hypothesized internal model representations of the aircraft dynamics, are constructed. The hypothesis testing scheme determines the relative probability that each pilot model best matches the DMS data. By observing the changes in probabilities, it is possible to determine when the pilot changes control strategy and which hypothesized pilot model best represent's the pilot's control behavior.

Broussard, J. R.

Statistical Hypothesis Testing in Wavelet Analysis: Theoretical Developments and Applications to Indian Rainfall

Statistical hypothesis tests in wavelet analysis are methods that assess the degree to which a wavelet quantity(e.g., power and coherence) exceeds background noise. Commonly, a point-wise approach is adopted in which a wavelet quantity at every point in a wavelet spectrum is individually compared to the critical level of the point-wise test. However, because adjacent wavelet coefficients are correlated and wavelet spectra often contain many wavelet quantities, the point-wise test can produce many false positive results that occur in clusters or patches. To circumvent the point-wise test drawbacks, it is necessary to implement the recently developed area-wise, geometric, cumulative area-wise, and topological significance tests, which are reviewed and developed in this paper. To improve the computational efficiency of the cumulative area-wise test, a simplified version of the testing procedure is created based on the idea that its output is the mean of individual estimates of statistical significance calculated from the geometric test applied at a set of point-wise significance levels. Ideal examples are used to show that the geometric and cumulative area-wise tests are unable to differentiate wavelet spectral features arising from singularity-like structures from those associated with periodicities. A cumulative arc-wise test is therefore developed to strictly test for periodicities by using normalized arc length, which is defined as the number of points composing a cross section of a patch divided by the wavelet scale in question. A previously proposed topological significance test is formalized using persistent homology profiles (PHPs) measuring the number of patches and holes corresponding to the set of all point-wise significance values. Ideal examples show that the PHPs can be used to distinguish time series containing signal components from those that are purely noise. To demonstrate the practical uses of the existing and newly developed statistical methodologies, a first comprehensive wavelet analysis of Indian rainfall is also provided. An R software package has been written by the author to implement the various testing procedures.

Wind speed

Approaches to vegetation mapping and ecophysiological hypothesis testing using combined information from TIMS, AVIRIS, and AIRSAR

The Tropical Rainforest Ecology Experiment (TREE) had two primary objectives: (1) to design a method for mapping vegetation in tropical regions using remote sensing and determine whether the result improves on available vegetation maps; and (2) to test a specific hypothesis on plant/water relations. Both objectives were thought achievable with the combined information from the Thermal Infrared Multispectral Scanner (TIMS), Airborne Visible/Infrared Imaging Spectrometer (AVIRIS), and Airborne Synthetic Aperture Radar (AIRSAR). Implicitly, two additional objectives were: (1) to ascertain that the range within each variable potentially measurable with the three instruments is large enough in the site, relative to the sensitivity of the instruments, so that differences between ecological groups may be detectable; and (2) to determine the ability of the three systems to quantify different variables and sensitivities. We found that the ranges in values of foliar nitrogen concentration, water availability, stand structure and species composition, and plant/water relations were large, even within the upland broadleaf vegetation type. The range was larger when other vegetation types were considered. Unfortunately, cloud cover and navigation errors compromised the utility of the TIMS and AVIRIS data. Nevertheless, the AIRSAR data alone appear to have improved on the available vegetation map for the study area. An example from an area converted to a farm is given to demonstrate how the combined information from AIRSAR, TIMS, and AVIRIS can uniquely identify distinct classes of land use. The example alludes to the potential utility of the three instruments for identifying vegetation at an ecological scale finer than vegetation types.

Oren, R.

Implications for high speed research: The relationship between sonic boom signature distortion and atmospheric turbulence

In this study there were two primary tasks. The first was to develop an algorithm for quantifying the distortion in a sonic boom. Such an algorithm should be somewhat automatic, with minimal human intervention. Once the algorithm was developed, it was used to test the hypothesis that the cause of a sonic boom distortion was due to atmospheric turbulence. This hypothesis testing was the second task. Using readily available sonic boom data, we statistically tested whether there was a correlation between the sonic boom distortion and the distance a boom traveled through atmospheric turbulence.

Sparrow, Victor W.

Design of a Uranium Dioxide Spheroidization System

The plasma spheroidization system (PSS) is the first process in the development of tungsten-uranium dioxide (W-UO2) fuel cermets. The PSS process improves particle spherocity and surface morphology for coating by chemical vapor deposition (CVD) process. Angular fully dense particles melt in an argon-hydrogen plasma jet at between 32-36 kW, and become spherical due to surface tension. Surrogate CeO2 powder was used in place of UO2 for system and process parameter development. Particles range in size from 100 - 50 microns in diameter. Student s t-test and hypothesis testing of two proportions statistical methods were applied to characterize and compare the spherocity of pre and post process powders. Particle spherocity was determined by irregularity parameter. Processed powders show great than 800% increase in the number of spherical particles over the stock powder with the mean spherocity only mildly improved. It is recommended that powders be processed two-three times in order to reach the desired spherocity, and that process parameters be optimized for a more narrow particles size range. Keywords: spherocity, spheroidization, plasma, uranium-dioxide, cermet, nuclear, propulsion

Cavender, Daniel P.

On Restructurable Control System Theory

The state of stochastic system and control theory as it impacts restructurable control issues is addressed. The multivariable characteristics of the control problem are addressed. The failure detection/identification problem is discussed as a multi-hypothesis testing problem. Control strategy reconfiguration, static multivariable controls, static failure hypothesis testing, dynamic multivariable controls, fault-tolerant control theory, dynamic hypothesis testing, generalized likelihood ratio (GLR) methods, and adaptive control are discussed.

Athans, M.

Statistical Analysis of Model Data for Operational Space Launch Weather Support at Kennedy Space Center and Cape Canaveral Air Force Station

The 12-km resolution North American Mesoscale (NAM) model (MesoNAM) is used by the 45th Weather Squadron (45 WS) Launch Weather Officers at Kennedy Space Center (KSC) and Cape Canaveral Air Force Station (CCAFS) to support space launch weather operations. The 45 WS tasked the Applied Meteorology Unit to conduct an objective statistics-based analysis of MesoNAM output compared to wind tower mesonet observations and then develop a an operational tool to display the results. The National Centers for Environmental Prediction began running the current version of the MesoNAM in mid-August 2006. The period of record for the dataset was 1 September 2006 - 31 January 2010. The AMU evaluated MesoNAM hourly forecasts from 0 to 84 hours based on model initialization times of 00, 06, 12 and 18 UTC. The MesoNAM forecast winds, temperature and dew point were compared to the observed values of these parameters from the sensors in the KSC/CCAFS wind tower network. The data sets were stratified by model initialization time, month and onshore/offshore flow for each wind tower. Statistics computed included bias (mean difference), standard deviation of the bias, root mean square error (RMSE) and a hypothesis test for bias = O. Twelve wind towers located in close proximity to key launch complexes were used for the statistical analysis with the sensors on the towers positioned at varying heights to include 6 ft, 30 ft, 54 ft, 60 ft, 90 ft, 162 ft, 204 ft and 230 ft depending on the launch vehicle and associated weather launch commit criteria being evaluated. These twelve wind towers support activities for the Space Shuttle (launch and landing), Delta IV, Atlas V and Falcon 9 launch vehicles. For all twelve towers, the results indicate a diurnal signal in the bias of temperature (T) and weaker but discernable diurnal signal in the bias of dewpoint temperature (T(sub d)) in the MesoNAM forecasts. Also, the standard deviation of the bias and RMSE of T, T(sub d), wind speed and wind direction indicated the model error increased with the forecast period all four parameters. The hypothesis testing uses statistics to determine the probability that a given hypothesis is true. The goal of using the hypothesis test was to determine if the model bias of any of the parameters assessed throughout the model forecast period was statistically zero. For th is dataset, if this test produced a value >= -1 .96 or <= 1.96 for a data point, then the bias at that point was effectively zero and the model forecast for that point was considered to have no error. A graphical user interface (GUI) was developed so the 45 WS would have an operational tool at their disposal that would be easy to navigate among the multiple stratifications of information to include tower locations, month, model initialization times, sensor heights and onshore/offshore flow. The AMU developed the GUI using HyperText Markup Language (HTML) so the tool could be used in most popular web browsers with computers running different operating systems such as Microsoft Windows and Linux.

Bauman, William H., III

Setting the Bar for the Replacement of the Probability of Collision Metric

To date, satellite conjunction assessment (CA) risk analysis has largely embraced the probability of collision (Pc) as the omnibus metric to evaluate collision likelihood, and its use in such assessments has mostly been straightforward: at the point at which a conjunction mitigation decision is required, the calculated Pc is compared to a threshold; and if the calculated Pc exceeds that threshold, then a mitigation action is warranted. With only minor variation, this approach is employed by major CA risk assessment centers (e.g., NASA, EUSST, CNES, JAXA) and is advanced as the preferred method in the published CA best practices handbooks. Despite this near unanimity of operational practice, there is a major strain of secondary literature critical of the Pc and willing to propose alternatives. Alfano (2005) pointed out the ability of the Pc to underrepresent the risk in certain situations and counselled a maximum Pc construct. Carpenter (2017, 2019) reiterated this criticism and proposed using instead a confidence interval on the miss distance. Balch et al. (2019) identified what they argued was a defect in the entire Bayesian Pc construct and believed that the use of a more conservative methodology based on covariance ellipsoid overlap was necessary. Delande (2022) introduced the framework of collision “plausibility” to the risk assessment process and sketched out how this might be used operationally. Elkantassi (2022) published a full development of the miss distance confidence interval approach and applied it to several worked examples. While these different approaches to collision risk assessment do differ in their details, they all converge on two central points: first, the Pc’s failure to give an adequate expression of the risk in dilution region situations is a fatal flaw; and second, a conjunction should be presumed risky and in need of mitigation until the evidence of the situation can establish otherwise. These criticisms, if correct, would counsel a number of modifications to current CA operational practice; as such, they force a re-examination of fundamental aspects of the CA problem, including the following: 1. Is the CA risk assessment a probability problem, a statistics problem, or something else? 2. If it is a statistics problem, does it lend itself naturally to a hypothesis test construction? 3. If it can be construed as a hypothesis test, what form should the null hypothesis take, to wit: what constraints exist on the choice of the null hypothesis, what selections are in best alignment with all of the attendant parameters of the problem, and what is implied philosophically by different choices? 4. What are the implications of using the different proposed risk assessment parameters for CA? This question should be answered both in determining how frequently the dilution region situation cited by the critics of the Pc actually appears in an operationally significant manner and the missed detection and false alarm rates of all of the proposed risk assessment metrics, compared both to the Pc and to each other. This paper explores and offers preliminary answers to the above questions, presenting a researched treatment of the philosophical nature of the CA problem and the null hypothesis choice that achieves the greatest consistency with all of the different aspects of operational CA conduct. It then profiles all of the different proposed risk assessment metrics enumerated in the earlier paragraph against an extremely large database of conjunction events at both the 550km and 700km altitudes. The combination of the philosophical exploration of the CA problem and the results of the profiling activity articulates what a risk assessment metric will need to demonstrate, in terms of both innate construction and performance, in order to be a true competitor to the Pc.

conjunction assessment

Summary and Annotated Bibliography of Measurement Error Corrections with Potential Application in Future Quesst Mission Community Noise Studies

This document is motivated by likely needs of the Quesst mission community response tests, which will culminate in data collection and estimation of dose-response regression relationships for consideration by domestic and international aviation regulators. Furthermore, basic research questions evaluating interactions between rates of community annoyance, dose levels, and indicators of the presence of rattle, vibration, and startle hinge on hypothesis testing in the context of regression models. For a variety of reasons, noise doses may be known only imprecisely and may not reflect the actual level experienced by responding subjects. These differences between true dose and estimated dose, be they systematic or random, constitute covariate measurement error. Available statistics literature speaks to the impacts of measurement error on regression models, both in terms of bias in estimated coefficients and predicted values, and in terms of the loss of statistical power for hypothesis testing. Given the particulars of a categorical annoyance response variable and a continuous noise dose predictor variable subject to measurement error during testing, the emphasis of this report is on findings and methods pertinent to generalized linear (and mixed) models likely to be employed during the Quesst mission community tests. We reach the following conclusions: 1. Of four reviewed methods, structural Bayesian measurement error models and simulation extrapolation (SIMEX) may be the most readily applicable to Quesst mission community noise study objectives. 2. If warranted, a linear measurement model can help model systematic sources of measurement error that the classical measurement error does not. 3. For its ready implementation and small additional input requirements, simulation extrapolation may be ideally suited for addressing secondary research questions involving interactions between annoyance, noise dose, and other factors through hypothesis testing. 4. For their flexibility and ability to propagate uncertainty, structural Bayesian hierarchical models have great appeal for mission purposes; some care may be needed in developing appropriate probability models describing actual noise exposure during testing. An annotated bibliography logs additional papers and resources that may be of value to analysts in other projects and disciplines.

Dose-Response Model

OTD Observations of Continental US Ground Flashes Detected by NLDN

Lightning optical flash parameters (e.g., radiance, area, duration, number of optical groups, and number of optical events) derived from almost 5 yrs of Optical Transient Detector (OTD) data are compared with peak current and multiplicity observations derived from the US National Lightning Detection Networkm (NLDN). Despite the relatively low lightning geolocation accuracy afforded by OTD, a total of 48,870 NLDN cloud-to-ground (CG) flashes were correlated with OTD flashes, or about 10,000 CGs per year. The median values of the above OTD flash parameters for the 48,870 CGs were, respectively: 0.137 J/square meters/sr/micrometers, 313.7 square kilometers, 0.189 s, 4 optical groups per CG, and 8 optical events per CG. Invoking the multiplicity data, the median number of optical groups per stroke was 2.5, and the median number of optical events per stroke was 5.0. Median values of peak current for negative and positive CGs were -21.6 kA and 17.8 kA, respectively, and as expected, the negative CGs had a larger average multiplicity than the positive CGs. A statistical summary is provided for all CGs, for positive and negative CGs, and for CGs from different seasons. Standard two-distribution hypothesis tests were perfonned to intercompare the population means of the various lightning parameters. In particular, and to greater than the 99% confidence level, it was found that positive CGs are on average more radiant, of greater areal extent, and are longer lasting than negative CGs. Rankings from a complete set of hypothesis tests between CGs of different polarities and from different seasons are also provided. Most notably, wintertime positive CGs tend to be more radiant, of greater areal extent, and longer lasting than any other group of CGs (i.e., negative springtime CGs, positive summertime CGs, etc.).

Koshak, William J.

Study Design to Test the Hypothesis That Long-Term Space Travel Harms the Human and Animal Immune Systems

The potential threat of immunosuppression and abnormal inflammatory responses in long-term space travel, leading to unusual predilection for opportunistic infections, malignancy, and death, is of ma or concern to the National Aeronautics and Space Administration (NASA) Program. This application has been devised to seek answers to questions of altered immunity in space travel raised by previous investigations spanning 30-plus years. We propose to do this with the help of knowledge gained by the discovery of the molecular basis of many primary and secondary immunodeficiency diseases and by application of molecular and genetic technology not previously available. Two areas of immunity that previously received little attention in space travel research will be emphasized: specific antibody responses and non-specific inflammation and adhesion. Both of these areas of research will not only add to the growing body of information on the potential effects of space travel on the immune system, but be able to delineate any functional alterations in systems important for antigen presentation, specific immune memory, and cell:cell and cell:endothelium interactions. By more precisely defining molecular dysfunction of components of the immune system, it is hoped that targeted methods of prevention of immune damage in space could be devised.

Shearer, William T.

Testing the Hypothesis of Young Martian Volcanism: Studies of the Tharsis Volcanoes and Adjacent Lava Plains

We experienced much success in reaching our stated goals in our original MDAP proposal. Our work made substantial contributions towards an integrated understanding of the counting and calibration of crater data on Mars, and changing nature of the Martian surface influenced by craters, water, and wind, and their general relationship to Martian geothermal history. We accomplished this while being to responsive to the rapid changes in the field brought about by several key NASA missions that returned data during the life of the grant. Our integrated effort included three stages: The first major area of research (Crater Count Research) was conducted by Jennifer Grier (P.I.), Lazslo Keszthelyi (Collaborator), William Hartmann (Collaborator), with assistance from Dan Berman (Graduate student) and concerned the mapping and the collection of crater count data on various Martian terrains. The second major area of study (Absolute Age Calibration) was conducted by William Bottke (Co-I) at SWRI, and concerned constraining the nature of the Moon and Mars impactor populations to create better absolute age calibrations for counted areas. The third major area of study was the integration and leverage of this effort with ongoing related Mars crater work at PSI (Integrated and Continuing Studies - Older Volcanoes), headed by David Crown (PSI Scientist), assisted by Les Bleamaster (PSI Scientist) and Dan Berman (Graduate Student).

Grier, Jennifer A.

Mini Survey of SDSS [OIII] AGN with Swift: Testing the Hypothesis that L(sub [OIII]) Traces AGN Luminosity

The number of AGN and their luminosity distribution are crucial parameters for our understanding of the AGN phenomenon. Recent work strongly suggests every massive galaxy has a central black hole. However most of these objects either are not radiating or have been very difficult to detect We are now in the era of large surveys, and the luminosity function (LF] of AGN has been estimated in various ways. In the X-ray band. Chandra and XMM surveys have revealed that the LF of hard X-ray selected AGN shows a strong luminosity-dependent evolution with a dramatic break towards low L(sub x) (at all z). This is seen for all types of AGN, but is stronger for the broad-line objects. In sharp contrast, the local LF of optically-selected samples shows no such break and no differences between narrow and broad-line objects. If as been suggested, hard X ray and optical emission line can both can be fair indicators of AGN activity, it is important to first understand how reliable these characteristics are if we hope to understand the apparent discrepancy in the LFs.

Source record