Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “hypothesis tests”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Hydrogen underground storage for grid electricity storage: An optimization study on techno-economic analysis

Here, this study performs a techno-economic analysis of hydrogen underground storage systems for grid electricity storage, evaluating their economic viability at the plant scale using dynamic optimization. It explores the feasibility of various system configurations and revenue models in the context of volatile electricity prices and the necessity for multiple revenue streams. The hypothesis tested is that large-scale hydrogen storage, despite its low round-trip efficiency, can be economically viable with the right mix of revenue streams. This study uses scenario-based analysis to assess the impacts of different system configurations, including engaging in time-shifting arbitrage, ancillary service markets and blending hydrogen with natural gas. Results indicate potential annual net cash flows of up to $\$$1.5 million from ancillary services integration and $\$$5.2 million from natural gas blending, contingent on specific system sizes. The study concludes that hydrogen underground storage for grid electricity storage can be profitable, and emphasizes that proper system design and precise electricity price forecasting are crucial for optimizing system performance and economic returns. This research sets the stage for further investigations into the scalability of hydrogen storage systems and their broader implications for grid electricity storage and energy market dynamics.

25 ENERGY STORAGE↗

Dynamic and single cell characterization of a CRISPR-interference toolset in Pseudomonas putida KT2440 for β-ketoadipate production from p -coumarate

We report Pseudomonas putida KT2440 is a well-studied bacterium for the conversion of lignin-derived aromatic compounds to bioproducts. The development of advanced genetic tools in P. putida has reduced the turnaround time for hypothesis testing and enabled the construction of strains capable of producing various products of interest. Here, we evaluate an inducible CRISPR-interference (CRISPRi) toolset on fluorescent, essential, and metabolic targets. Nuclease-deficient Cas9 (dCas9) expressed with the arabinose (8K)-inducible promoter was shown to be tightly regulated across various media conditions and when targeting essential genes. In addition to bulk growth data, single cell time lapse microscopy was conducted, which revealed intrinsic heterogeneity in knockdown rate within an isoclonal population. The dynamics of knockdown were studied across genomic targets in exponentially-growing cells, revealing a universal 1.75 ± 0.38 hour quiescent phase after induction where 1.5 ± 0.35 doublings occur before a phenotypic response is observed. To demonstrate application of this CRISPRi toolset, β-ketoadipate, a monomer for performance-advantaged nylon, was produced at a 4.39 ± 0.5 g/L and yield of 0.76 ± 0.10 mol/mol from p-coumarate, a hydroxycinnamic acid that can be derived from grasses. These cultivation metrics were achieved by using the higher strength IPTG (1K)-inducible promoter to knockdown the pcaIJ operon in the βKA pathway during early exponential phase. This allowed the majority of the carbon to be shunted into the desired product while eliminating the need for a supplemental carbon and energy source to support growth and maintenance.

59 BASIC BIOLOGICAL SCIENCES↗

Multivariable degradation modeling and life prediction using multivariate fractional Brownian motion

In system prognostics and health management, multivariable degradation models have been widely developed to predict the life of complex systems using degradation data of multiple Performance Characteristics (PCs). Recent studies have detected a Long-Term Memory (LTM) effect among the degradation process of various PCs, implying a strong coupling phenomenon between the future degradation behavior and historical degradation trajectory. Although the LTM has been widely integrated into single-PC-based degradation modeling, it has not been considered in multi-PC-based scenarios. To capture LTM among multiple PCs, this article proposes a novel LTM-integrated Multivariate Degradation Model (MDM) for system life prediction based on multivariate fractional Brownian motion, which simultaneously incorporates the cross-correlation among different PCs. To estimate parameters of the LTM-integrated MDM, a maximum likelihood method is developed. Here, two likelihood-ratio hypothesis tests are developed to test the existence of the overall and individual LTM effect among multiple PCs. Both simulation studies and physical experiments on the performance degradation of solar energy conversion and storage devices are conducted to validate the proposed model. Results reveal that the proposed LTM-integrated MDM significantly outperforms existing MDMs in life prediction, while the lifetime uncertainty is heavily underestimated by those traditional approaches that neglect the LTM.

42 ENGINEERING↗

A Mass‐Conserving‐Perceptron for Machine‐Learning‐Based Modeling of Geoscientific Systems

Although decades of effort have been devoted to building Physical-Conceptual (PC) models for predicting the time-series evolution of geoscientific systems, recent work shows that Machine Learning (ML) based Gated Recurrent Neural Network technology can be used to develop models that are much more accurate. However, the difficulty of extracting physical understanding from ML-based models complicates their utility for enhancing scientific knowledge regarding system structure and function. Here, we propose a physically interpretable Mass-Conserving-Perceptron (MCP) as a way to bridge the gap between PC-based and ML-based modeling approaches. The MCP exploits the inherent isomorphism between the directed graph structures underlying both PC models and GRNNs to explicitly represent the mass-conserving nature of physical processes while enabling the functional nature of such processes to be directly learned (in an interpretable manner) from available data using off-the-shelf ML technology. As a proof of concept, we investigate the functional expressivity (capacity) of the MCP, explore its ability to parsimoniously represent the rainfall-runoff (RR) dynamics of the Leaf River Basin, and demonstrate its utility for scientific hypothesis testing. To conclude, we discuss extensions of the concept to enable ML-based physical-conceptual representation of the coupled nature of mass-energy-information flows through geoscientific systems.

58 GEOSCIENCES↗

On-the-fly autonomous control of neutron diffraction via physics-informed Bayesian active learning

We demonstrate the first live, autonomous control over neutron diffraction experiments by developing and deploying ANDiE: the autonomous neutron diffraction explorer. Neutron scattering is a unique and versatile characterization technique for probing the magnetic structure and behavior of materials. However, instruments at neutron scattering facilities in the world is limited, and instruments at such facilities are perennially oversubscribed. We demonstrate a significant reduction in experimental time required for neutron diffraction experiments by implementation of autonomous navigation of measurement parameter space through machine learning. Prior scientific knowledge and Bayesian active learning are used to dynamically steer the sequence of measurements. We show that ANDiE can experimentally determine the magnetic ordering transition of both MnO and Fe 1.09 Te all while providing a fivefold enhancement in measurement efficiency. Furthermore, in a hypothesis testing post-processing step, ANDiE can determine transition behavior from a set of possible physical models. ANDiE's active learning approach is broadly applicable to a variety of neutron-based experiments and can open the door for neutron scattering as a tool of accelerated materials discovery.

36 MATERIALS SCIENCE↗

Large-database cross-verification and validation of tokamak transport models using baselines for comparison

State-of-the-art 1D transport solvers ASTRA and TRANSP are verified, then validated across a large database of semi-randomly selected, time-dependent DIII-D discharges. Various empirical models are provided as baselines to contextualize the validation figures of merit using statistical hypothesis tests. For predicting plasma temperature profiles, no statistically significant advantage is found for the ASTRA and TRANSP simulators over a baseline empirical (two-parameter) model. For predicting stored energy, a significant advantage is found for the simulators over a baseline empirical model based on confinement time scaling. Uncertainty in the results due to diagnostic and profile fitting uncertainties is approximated and determined to be insignificant due in part to the large quantity of discharges employed in the study. Advantages are discussed for validation methodologies like this one that employ (1) large databases and (2) baselines for comparison that are specific to the intended use-case of the model.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Not even 6 dB: Gaussian quantum illumination in thermal background

Abstract In analyses of target detection with Gaussian state transmitters in a thermal background, the thermal occupation is taken to depend on the target reflectivity in a way which simplifies the analysis of the symmetric quantum hypothesis testing problem. However, this assumption precludes comparison of target detection performance between an arbitrary transmitter and a vacuum state transmitter, i.e. ‘detection without illumination’, which is relevant in a bright thermal background because a target can be detected by its optical shadow or some other perturbation of the background. Using a target-agnostic thermal environment leads to the result that the oft-claimed 6 dB possible reduction in the quantum Chernoff exponent for a two-mode squeezed vacuum transmitter over a coherent state transmitter in high-occupation thermal background is an unachievable limiting value, only occurring in a limit in which the target detection problem is ill-posed. Further analyzing quantum illumination in a target-agnostic thermal environment shows that a weak single-mode squeezed transmitter performs worse than ‘no illumination’, which is explained by the noise-increasing property of reflected low-intensity squeezed light.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Reductive quantum phase estimation

Estimating a quantum phase is a necessary task in a wide range of fields of quantum science. To accomplish this task, two well-known methods have been developed in distinct contexts, namely, Ramsey interferometry (RI) in atomic and molecular physics and quantum phase estimation (QPE) in quantum computing. We demonstrate that these canonical examples are instances of a larger class of phase estimation protocols, which we call reductive quantum phase estimation (RQPE) circuits. Here, we present an explicit algorithm that allows one to create an RQPE circuit. This circuit distinguishes an arbitrary set of phases with a smaller number of qubits and unitary applications, thereby solving a general class of quantum hypothesis testing to which RI and QPE belong. We further demonstrate a tradeoff between measurement precision and phase distinguishability, which allows one to tune the circuit to be optimal for a specific application. Published by the American Physical Society 2024

Papadopoulos, Nicholas J. C. (ORCID:00000002635700↗

High Impedance Fault Detection Through Quasi-Static State Estimation: A Parameter Error Modeling Approach

This paper presents a model for detecting high impedance faults using parameter error modeling and a two step per-phase weighted-least squares state estimation process. The proposed scheme leverages the use of Phasor Measurement Units and synthetic measurements to identify per-phase power flow and injection measurements which indicate a parameter error through ?2 Hypothesis Testing applied to the composed measurement error. Although current and voltage waveforms are commonly analyzed for high-impedance fault detection, wide area power flow and injection measurements, which are already inherent to the state estimation process, also show promise for real-world high-impedance fault detection applications. The error distributions after detection share the measurement function error spread observed in proven parameter error diagnostics and can be applied to high-impedance fault identification. Further, this error spread across measurement functions related to the fault will be clearly discerned from measurement error. Case studies are performed on the IEEE 33-Bus Distribution System along with the proposed model in Simulink.

Cooper, Austin↗

Trust Model System for the Energy Grid of Things Network Communications

Network communication is crucial in the Energy Grid of Things (EGoT). Without a network connection, the energy grid becomes just a power grid where the energy resources are available to the customer uni-directionally. A mechanism to analyze and optimize the energy usage of the grid can only happen through a medium, a communications network, that enables information exchange between the grid participants and the service provider. Security implementers of EGoT network communication take extraordinary measures to ensure the safety of the energy grid, a critical infrastructure, as well as the safety and privacy of the grid participants. With the dynamic nature of network communication of the EGoT, the information provided by the customer or the service provider can be falsified by a malicious attacker. Therefore, a trust model is necessary to monitor any abnormal activities. This paper describes a distributed trust model system that meets the need of the EGoT. This paper describes methods for evaluating and improving the distributed trust model using standard hypothesis testing metrics such as true positive, false positive, true negative, false negative, equal error rate, and F1 score. Example calculations are shown based on generated sample data.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Trust Model Measurements for the Energy Grid of Things

Information security is essential for the reliable operation of an Energy Grid of Things (EGoT). In addition to basic information security protocols as defined by published standards, there is a need for a monitoring function that measures the trustworthiness of the various actors participating in an EGoT. We describe in this paper the implementation and evaluation of a Distributed Trust Model that was developed specifically for monitoring communication within an EGoT. We then show how the model parameters are set using statistical measures for hypothesis testing.

Energy Grid of Things, EGoT, Smart Grid Security, ↗

A Contextually Supervised Optimization-Based HVAC Load Disaggregation Methodology

This paper presents a novel contextually supervised optimization-based approach for disaggregating heating, ventilation, and air-conditioning (HVAC) loads using smart meter or Supervisory Control and Data Acquisition data. To disaggregate the load into HVAC loads, large and infrequently used loads (LIUL), and base loads, we formulate an optimization problem to minimize a set of five loss terms, consisting of the reconstruction errors of the overall load profile, the ramp rate losses, and three distinct loss functions linked with the HVAC load, base load, and LIUL, respectively. To enhance accuracy, we incorporate two forms of contextual information into the problem formulation. First, we utilize mutual information to estimate HVAC energy consumption. Second, we employ a base load dictionary to constrain HVAC load estimation errors. The obtained HVAC load profiles are fine-tuned by abnormal ramp detection followed by binary hypothesis testing. Here, the proposed method is developed and tested using sub-metered residential and commercial building data. Simulation results show that the proposed method outperforms existing methods across various data resolutions and load aggregation levels, showing excellent transferability and generalizability.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Covariate Dependent Sparse Functional Data Analysis

This study proposes a method to incorporate covariate information into sparse functional data analysis. The method aims at cases where each subject has a limited number of longitudinal measurements and is associated with static covariates. This research is motivated by several use cases in practice. One representative example is void swelling, a nuclear-specific material degradation mechanism. Void swelling is affected by many covariates, including alloy composition and irradiation type. How to accurately model the complicated joint effects of such covariates on the swelling process is the key to mitigating the effect of swelling and ensuring safe operation. Unlike most of the existing methods, the proposed method can handle high-dimensional covariates with the informative covariate identification procedure and sparse and irregularly spaced measurements, that is, does not require complete or dense observations. The main innovation of the proposed method is that we model the variation coming from covariates and the variation left conditioned on covariates, such that the functional principal component analysis and Gaussian process can be conducted in a unified manner. Further, we also propose a systematic approach to identify important covariates in the hypothesis testing context. The methodology is demonstrated on applications in nuclear engineering and healthcare and simulation studies.

42 ENGINEERING↗

Model Inputs, Outputs, and Scripts associated with: “Spatial microbial respiration variations in the hyporheic zones within the Columbia River Basin”

This data package is associated with the publication “Spatial microbial respiration variations in the hyporheic zones within the Columbia River Basin” published in the Journal of Geophysical Research: Biogeosciences (Son et al. 2022) available at doi: 10.1029/2021JG006654. This data package includes the key model inputs/outputs of the river corridor model for the Columbia River Basin (CRB) and the model source codes, which were used in the manuscript. The model is a carbon-nitrogen-coupled river corridor model (RCM), and the model is used to quantify hyporheic zone (HZ) aerobic and anaerobic respiration at the NHDPLUS stream reach scales. The RCM used in this study combines empirical substrate models derived from observations and three microbially driven reactions to compute respiration of the HZ for each National Hydrography Dataset (NHD) reach within the CRB. The reactions in HZs of each NHD reach include anaerobic respiration and two-step anaerobic respiration via denitrification. Our HZ respiration estimates are limited to the lotic (or flowing) stream/river systems, and do not account for the respiration process in water column. Note that the RCM only simulates the HZ’s contribution to the dissolved carbon dioxide (CO2) concentrations in the streams, and the CO2 emissions to the atmosphere are not modelled. The model computes at hourly timesteps because of the fast reaction rates. The key input data of the model are exchange flux, residence time, and stream solute (dissolved organic carbon (DOC), dissolved oxygen (DO), and nitrate concentrations). These inputs are constant over time and represent long-term averaged values.This modeling framework successfully quantified HZ respiration components over multiple scales. It revealed key mechanisms driving the spatial variation of HZ aerobic and anaerobic respiration in reaches with varying hydrologic and substrate conditions. Thus, this modeling study offers a testing hypothesis in different river system (e.g., climate and biomes) for the HZ respiration processes, and can be used as a sampling design tool for large-scale HZ experimental studies.This dataset contains five folders: (1) model_inputs, (2) model_outputs, (3) Rscripts, (4) figures, and (5) model_codes. It also contains a readme, file level metadata (FLMD), and data dictionary (dd). Please see the FLMD for a list of all the files contained in this data package and descriptions for each. The model_inputs folder contains the model inputs used to drive the model simulations. The model_outputs folder contains key model output files from the river corridor model. The Rscripts folder contains the Rscripts for pre- and post- processing model results. The figures folder contains the raw figures associated with the manuscript. The model_codes folder includes key model source codes/input files. All files are .jpg, .jpeg, .out, .e, .od, .dat, .sub, .F90, .0, .R, .sbx, .cpg, .sbn, .shx, .shp, .dbf, .prj, .tfw, .tif, .xml, .pdf, or .csv.

54 ENVIRONMENTAL SCIENCES↗

Forecasting generative amplification

Generative networks are perfect tools to enhance the speed and precision of LHC simulations. Especially when generating events beyond the size of the training dataset, it is important to understand their statistical precision. We present two complementary methods to estimate the amplification factor without large holdout datasets. Averaging amplification uses Bayesian networks or ensembling to estimate amplification from the precision of integrals over given phase-space volumes. Differential amplification uses hypothesis testing to quantify amplification without any resolution loss. Applied to state-of-the-art event generators, both methods indicate that amplification is already possible in specific regions of phase space.

Bahl, Henning [Heidelberg Univ. (Germany)] (ORCID:↗

Systems Analysis of the Physiological and Molecular Mechanisms of Sorghum Nitrogen Use Efficiency, Water Use Efficiency and Interactions with the Soil Microbiome (Final Report for DE-SC0014395)

The specific project objectives were to: 1) Conduct deep census surveys of root microbiomes concurrent with phenotypic characterizations of a diverse panel of sorghum genotypes across multiple years to define the microbes associated with the most productive lines under drought and low nitrogen conditions. 2) Associate systems-level genotypic, microbial, and environmental factors with improved sorghum performance using robust statistical approaches. 3) Develop culture collections of sorghum root/leaf associated microbes that recapitulate root-enriched sequences defined in the census. 4) Perform controlled environment experiments for in-depth characterization and hypothesis testing of G sorghum x G microbe x E interactions . Validate physiological mechanisms, map genetic loci for stress tolerance, and determine the persistence of optimal microbial strains under greenhouse and field conditions.

59 BASIC BIOLOGICAL SCIENCES↗

3P Program: Phenotyping X Prediction = Productivity (Final Scientific/Technical Report)

The goal of the 3P Program was to establish integrated, real-time phenotyping and to analyze above- and below-ground plant architecture and total carbon partitioning and allocation to predict heterosis and develop superior crop hybrids by fully leveraging the Sorghum gene pool. There were two overarching themes: 1) the development of a new crop improvement approach utilizing advances in high-throughput phenotyping (HTP), computing, and genomics for public dissemination and 2) leveraging this platform for sorghum crop improvement and commercialization. The Clemson team worked on creating genomic resources and using both statistical learning and high-throughput phenotyping in genomics-assisted breeding. Research was broadly interested in the genetics of carbon partitioning, with the aim of improving crop performance and achieving sustainability. The technology and resources created can be readily found in the public domain and serve to advance scientific understanding of crop genomics and breeding. Genomic prediction was able to identify top crosses to be made, and a hybrid prediction pipeline is in place to drive year-over-year genetic gain. Roots have long been ignored by plant breeders and agronomists, not because they are unimportant but because they are hard to measure. This is an untapped white space of potential insight and innovation. To address this, Hi Fidelity Genetics developed the RootTracker to measure roots in the field on a continuous basis. A database system called RootTracker Tracker was developed to handle data coming from the RootTrackers. In using this device, valuable data was observed for plant breeding, hydrochemical development, and other agricultural biology applications. Carnegie Mellon’s goal was developing new techniques to generate high-resolution 3D models of plants from data collected in the field. The idea was that more useful and more informative phenotypes could be extracted by resolving small features, such as seeds and flowers, and that by modeling in 3D, the spatial structure of plants could be examined. To achieve this, multiple images collected by a new small format structured light stereo imager were fused together. A sorghum panicle modeling pipeline was developed to allow the collection and processing of data. Carolina Seed Systems is an agricultural technology company focused on decarbonizing the agricultural system. Their technology pipeline serves to drive fundamental progress towards creation and distribution of carbon negative crops. The genomic and the engineering technology developed through the 3P Program was leveraged to deliver both value and sustainability from the grower to the consumer. Promising sorghum hybrids were scaled up and commercialized. The overall goal of our research was to integrate, create, and deploy genetic and engineering concepts and technologies to enhance crop productivity in a sustainable fashion. The combination of public and private partners allowed the basic research and hypothesis testing to be quickly accelerated for commercial application by the companies yet maintained that the core framework and academic insights remain in the public domain for continued market disruption, competition, and innovation.

59 BASIC BIOLOGICAL SCIENCES↗

Integration of Waveform Simulation Methods

The generation of synthetic seismograms through simulation is a fundamental tool of seismology required to run quantitative hypothesis tests. A variety of approaches have been developed throughout the seismological community and each has their own specific user interface based on their implementation. This causes a challenge to researchers who will need to learn new interfaces with each new software they wish to use and create substantial challenges when attempting to compare results from different tools. Here we provide a unified interface that facilitates interoperability amongst several simulation tools through a modern containerized Python package. Further, this package includes post-processing analysis modules designed to facilitate end-to-end analysis of synthetic seismograms. In this report we present the conceptual guidance and an example implementation of the new Waveform Simulation Framework.

58 GEOSCIENCES↗