Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Bayesian experimental design”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Spatial‐Uniformity–Driven Bayesian Optimization for Rapid Development of Printed Perovskite Solar Cells

Printed metal halide perovskites can enable rapid, roll-to-roll manufacturing of a broad class of optoelectronics—flexible solar cells and imagers among them—while promising cost and speed advantages over incumbent silicon. However, though current methods offer high throughput and patterning capabilities, perovskite films’ spatial heterogeneity remains a challenge for large-area devices. Here, a spatial-uniformity-driven Bayesian optimization (BO) approach is leveraged to accelerate the development of printed perovskite solar cells and improve large-area device performance. Using a BO surrogate model, a 6D design space of ink chemistry and printing physics is explored via extensive iterative experimentation (≈100) informed by an objective function capturing spatial photoluminescence (PL) variance. It is discovered that optimizing for uniformity drives rapid advances in photovoltaic performance, yielding ≈20% power conversion efficiency (PCE) for small area (0.134 cm 2 ) devices and > 16% for large area (1 cm 2 ) devices. This machine-learning approach simultaneously enables rheological comparison of ink formulations that accelerate the leveling of Saffman-Taylor artifacts and improve film uniformity. Here, this showcases uniformity-driven BO as an efficient approach for uncovering the key printing physics and mitigating spatial heterogeneity to enable device scaling beyond small cell areas.

14 SOLAR ENERGY↗

Fully Bayesian Analysis With Model Inadequacy Correction For Nuclear Graphite Property Models With Hierarchical Variance Structure

Nuclear-grade graphites are extensively utilized in the core designs of various advanced nuclear reactors. Within the reactor environment, graphite is subjected to prolonged exposure to extreme conditions, including high temperatures, radiation, and potentially molten salt and oxygen. Such exposure can induce several degradation mechanisms in graphite, such as nonuniform volumetric strains caused by irradiation and thermal expansion, leading to stresses that may compromise the performance of graphite components. Assessing component integrity, forecasting component performance over the reactor's lifespan, and developing design standards necessitate robust tools for predicting fracture initiation and propagation in graphite structural components within nuclear reactors. This code enables the Bayesian calibration of properties for nuclear-grade graphites. Using a hierarchical Bayesian approach, multiple experimental data sources are combined to develop Gaussian process models for the properties. Using the Kennedy O'Hagan framework, the uncertainties due inadequacies in the model and the inherent spread in the experimental data are quantified.

Dhulipala, Som Lakshmi NarasimhaLakshmi Narasimha ↗

Autonomous organic synthesis for redox flow batteries via flexible batch Bayesian optimization

Traditional trial-and-error methods for materials discovery are inefficient to meet the urgent demands posed by the rapid progression of climate change. This urgency has driven the increasing interest in integrating robotics and machine learning into materials research to accelerate experimental learning. However, idealized decision-making frameworks to achieve maximum sampling efficiency are not always compatible with high-throughput experimental workflows inside a laboratory. For multi-step chemical processes, differences in hardware capacities can complicate the digital framework by introducing constraints on the maximum number of samples in each step of the experiment, hence causing varying batch sizes in variable selection within the same batch. Therefore, designing flexible sampling algorithms is necessary to accommodate the multi-step synthesis with practical constraints unique to each high-throughput workflow. In this work, we designed and employed three strategies on a high-throughput robotic platform to optimize the sulfonation reaction of redox-active molecules used in flow batteries. Our strategies adapt to the multi-step experimental workflow, where their formulation and heating steps are separate, causing varying batch size requirements. By strategically sampling using clustering and mixed-variable batch Bayesian optimization, we were able to iteratively identify optimal conditions that maximize the yields. Our work presents a flexible approach that allows tailoring the machine learning decision-making to suit the practical constraints in individual high-throughput experimental platforms, followed by performing resource-efficient yield optimization using available open-source Python libraries.

Tamura, Clara [Univ. of Washington, Seattle, WA (U↗

Li-ion battery design through microstructural optimization using generative AI

Lithium-ion batteries are used across various applications, necessitating tailored cell designs to enhance performance. Optimizing electrode manufacturing parameters is a key route to achieving this, as these parameters directly influence the microstructure and performance of the cells. However, linking process parameters to performance is complex, and experimental or modeling campaigns are often slow and expensive. This study introduces a fast computational optimization framework for electrode manufacturing parameters. A generative model, trained on a small dataset of microstructural images associated with different manufacturing parameters, efficiently generates representative microstructures for new parameters. This model is integrated into a Bayesian optimization loop that includes microstructure generation, characterization, and simulation, aiming to find optimal manufacturing parameters for a particular application. Significant improvement in the energy density of a 4680 cell is achieved through bespoke cell design, highlighting the importance of cell-scale normalization. The framework’s modularity allows its application to various advanced materials manufacturing scenarios.

batteries↗

Resolving root causes of experiment discrepancies guided by machine learning

Abstract Scientists rely on accurate experimental data to explain nature and then harness this knowledge for applications addressing human needs. However, discrepancies between experiments of the same observable can impede scientific progress if one does not understand the underlying causes. Here, we developed a process that unravels data discrepancies by first using Bayesian machine learning to relate discrepancies to few of many, potentially biasing metadata features that encode experiment procedures. This machine learning output guides human experts to study discrepancy causes by simulating suspicious aspects of historical experiments or designing modern ones to address open questions. The study findings then lead to rejecting or correcting historical data on firm scientific bases. This process is demonstrated for the energy spectrum of neutrons emitted promptly (<1 ns) after fission of 252 Cf, a trusted nuclear physics Standard. It reduces the spread in experimental 252 Cf spectra by up to a factor of 6.

Neudecker, D. (ORCID:0000000339200627)↗

Bayesian batch optimization for molybdenum versus tungsten inertial confinement fusion double shell target design

Access to reliable, clean energy sources is a major concern for national security. Much research is focused on the “grand challenge” of producing energy via controlled fusion reactions in a laboratory setting. For fusion experiments, specifically inertial confinement fusion (ICF), to produce sufficient energy, the fusion reactions in the ICF fuel need to become self-sustaining and burn deuterium-tritium (DT) fuel efficiently. The recent record-breaking NIF ignition shot was able to achieve this goal as well as produce more energy than used to drive the experiment. This achievement brings self-sustaining fusion-based power systems closer than ever before, capable of providing humans with access to secure, renewable energy. In order to further progress toward the actualization of such power systems, more ICF experiments need to be conducted at large laser facilities such as the United States's National Ignition Facility (NIF) or France's Laser Mega-Joule. The high cost per shot and limited number of shots that are possible per year make it prohibitive to perform large numbers of experiments. As such, experimental design relies heavily on complex predictive physics simulations for high-fidelity “preshot” analysis. These multidimensional, multi-physics, high-fidelity simulations have to account for a variety of input parameters as well as modeling the extreme conditions (pressures and densities) present at ignition. Such simulations (especially in 3D) can become computationally prohibitive to turn around for each ICF experiment. In this work, we explore using Bayesian optimization with Gaussian processes (GPs) to find optimal designs for ICF double shell targets, while keeping computational costs to manageable levels. These double shell targets have an inner shell that grades from beryllium on the outer surface to the higher Z material molybdenum, as opposed to the nominally used tungsten, on the inside in order to trade off between the high performance associated with high density inner shells and capsule stability. We describe our results for “capsule-only” xRAGE simulations to study the physics between different capsule designs, inner shell materials, and potential for future experiments.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

PyOED: An Extensible Suite for Data Assimilation and Model-Constrained Optimal Design of Experiments

This article describes PyOED, a highly extensible scientific package that enables developing and testing model-constrained optimal experimental design (OED) for inverse problems. Specifically, PyOED aims to be a comprehensive Python toolkit for model-constrained OED. The package targets scientists and researchers interested in understanding the details of OED formulations and approaches. It is also meant to enable researchers to experiment with standard and innovative OED technologies with a wide range of test problems (e.g., simulation models). OED, inverse problems (e.g., Bayesian inversion), and data assimilation (DA) are closely related research fields, and their formulations overlap significantly. Thus, PyOED is continuously being expanded with a plethora of Bayesian inversion, DA, and OED methods as well as new scientific simulation models, observation error models, and observation operators. These pieces are added such that they can be permuted to enable testing OED methods in various settings of varying complexities. The PyOED core is completely written in Python and utilizes the inherent object-oriented capabilities; however, the current version of PyOED is meant to be extensible rather than scalable. Specifically, PyOED is developed to “enable rapid development and benchmarking of OED methods with minimal coding effort and to maximize code reutilization.” This article provides a brief description of the PyOED layout and philosophy and provides a set of exemplary test cases and tutorials to demonstrate the potential of the package.

97 MATHEMATICS AND COMPUTING↗

Assessing the design of integrated methane sensing networks

Abstract While methane is the second largest contributor to global warming after carbon dioxide, it has a larger warming effect over a much shorter lifetime. Despite accelerated technological efforts to radically reduce global carbon dioxide emissions, rapid reductions in methane emissions are needed to limit near-term warming. Being primarily emitted as a byproduct from agricultural activities and energy extraction, methane is currently monitored via bottom–up (i.e. activity level) or top–down (via airborne or satellite retrievals) approaches. However, significant methane leaks remain undetected and emission rates are challenging to characterize with current monitoring frameworks. In this paper, we study the design of a layered monitoring approach that combines bottom–up and top–down approaches as an integrated sensing network. By recognizing that varying meteorological conditions and emission rates impact the efficacy of bottom–up monitoring, we develop a probabilistic approach to optimal sensor placement in its bottom–up network. Subsequently, we derive an inverse Bayesian framework to quantify the improvement that a design-optimized integrated framework has on emission-rate quantifications and their uncertainties. We find that under realistic meteorological conditions, the overall error in estimating the true emission rates is approximately 1.3 times higher, with their uncertainties being approximately 2.4 times higher, when using a randomized network over an optimized network, highlighting the importance of optimizing the design of integrated methane sensing networks. Further, we find that optimized networks can improve scenario coverage fractions by more than a factor of 2 over experimentally-studied networks, and identify a budget threshold beyond which the rate of optimized-network coverage improvement exhibits diminishing returns, suggesting that strategic sensor placement is also crucial for maximizing network efficiency.

54 ENVIRONMENTAL SCIENCES↗

Kinetics Modeling and Reactor Design Study of Glucose-to-Terpenes Cell-Free Conversion

Cell-free systems offer many advantages over traditional biological conversion by eliminating biological growth constraints. It also offers easy manipulation and finetuning of the reaction conditions for each individual enzyme. The conversion of cellulosic glucose to Limonene, a terpene, is a promising pathway for producing fuels and chemicals. Recent advances in developing cell-free systems focuses on bench scale optimization of terpene yield and to demonstrate its feasibility towards commercialization [1,2]. There is significant knowledge gap regarding reaction kinetics of these cell-free systems to further study how it will perform at larger scale. We present here, our studies on reaction kinetics and reactor design implications of cell-free glucose to Limonene conversion to facilitate the further development and commercialization of this process. We developed a novel kinetic model based on the metabolic-network structure of the cell-free system with multi-substrate reversible Michaelis-Menten rate law. To estimate kinetic parameters for this system of rate equations, we employed Bayesian optimization to perform global search with the assistance of gaussian processes to balance exploration and exploitation. The model parameters estimated showed good results compared with experimental data. The estimated parameters were used to perform sensitivity analysis. We found that Hexokinase is one of the most critical enzymes that affect the conversion of the glucose. We also observed that abundance of co-factors is also critical to the conversion of glucose to limonene. We investigated packed bed reactors with enzymes immobilized on the surface of particles to convert glucose stream into Limonene for larger scale production. The reactor design such as particle size, enzyme loading, and flow rate are found to be critical for improving yields. [1] Dudley, Q.M., Nash, C.J. and Jewett, M.C., 2019. Synthetic Biology, 4(1), p.ysz003. [2] Korman, T.P., Opgenorth, P.H. and Bowie, J.U., 2017. Nature communications, 8(1), p.15526.

09 BIOMASS FUELS↗

Superlative mechanical energy absorbing efficiency discovered through self-driving lab-human partnership

Energy absorbing efficiency is a key determinant of a structure’s ability to provide mechanical protection and is defined by the amount of energy that can be absorbed prior to stresses increasing to a level that damages the system to be protected. Here, we explore the energy absorbing efficiency of additively manufactured polymer structures by using a self-driving lab (SDL) to perform >25,000 physical experiments on generalized cylindrical shells. We use a human-SDL collaborative approach where experiments are selected from over trillions of candidates in an 11-dimensional parameter space using Bayesian optimization and then automatically performed while the human team monitors progress to periodically modify aspects of the system. The result of this human-SDL campaign is the discovery of a structure with a 75.2% energy absorbing efficiency and a library of experimental data that reveals transferable principles for designing tough structures.

42 ENGINEERING↗

Data‐Driven Engineering of Thermostable Collagen‐Mimetic Peptoid Triple Helices

Collagen-mimetic peptides (CMPs) are engineered molecules designed to replicate the triple-helical structure of natural collagen. A repeating x–y-Gly sequence is the defining motif of CMPs and is critical to their triple-helical structure and stability. Substitutions to the residues occupying the x and y positions present a means to modulate the CMP structure and properties. Peptoid residues—N-substituted glycine derivatives—present an attractive potential substitution due to their thermal stability, proteolytic resistance, biocompatibility, and diverse palette of non-natural side chains, but also tend to introduce a high degree of backbone flexibility that can diminish the stability of the triple helix. In this work, we report a computational active learning cycle comprising molecular dynamics simulation, Gaussian process regression, and Bayesian optimization to computationally identify a number of promising peptoid substitutions predicted to stabilize the desired quaternary structure through side chain interactions and produce stable peptoid-based collagen-like triple helices. To experimentally test the computational predictions, a top candidate identified by the screen was synthesized and imaged using scanning electron microscopy to resolve fibril-like bundles consistent with collagen-like triple helices. This work predicts a number of CMP peptoid substitutions capable of forming stable triple-helical structures, presents a generalizable design strategy for engineering desired peptoid structures, and opens new avenues for the design of peptoid-based biomimetic materials.

active learning↗

Optimal experimental design: Formulations and computations

Questions of ‘how best to acquire data’ are essential to modelling and prediction in the natural and social sciences, engineering applications, and beyond. Optimal experimental design (OED) formalizes these questions and creates computational methods to answer them. This article presents a systematic survey of modern OED, from its foundations in classical design theory to current research involving OED for complex models. We begin by reviewing criteria used to formulate an OED problem and thus to encode the goal of performing an experiment. We emphasize the flexibility of the Bayesian and decision-theoretic approach, which encompasses information-based criteria that are well-suited to nonlinear and non-Gaussian statistical models. We then discuss methods for estimating or bounding the values of these design criteria; this endeavour can be quite challenging due to strong nonlinearities, high parameter dimension, large per-sample costs, or settings where the model is implicit. A complementary set of computational issues involves optimization methods used to find a design; we discuss such methods in the discrete (combinatorial) setting of observation selection and in settings where an exact design can be continuously parametrized. Finally we present emerging methods for sequential OED that build non-myopic design policies, rather than explicit designs; these methods naturally adapt to the outcomes of past experiments in proposing new experiments, while seeking coordination among all experiments to be performed. Throughout, we highlight important open questions and challenges.

97 MATHEMATICS AND COMPUTING↗

Investigating Kinetic Mechanisms of Soot Formation in Plasma Pyrolysis of Methane via Active Learning (Final Technical Report)

Plasma pyrolysis of methane is an effective route for zero-carbon hydrogen production. Yet, soot generated from pyrolysis of hydrocarbons is detrimental to the climate and human health. There is ample experimental and theoretical evidence that suggests polycyclic aromatic hydrocarbons (PAHs) are the molecular precursors to soot particles. The reaction pathways of PAH formation are intricately dependent on a multitude of process parameters, whose kinetic mechanisms are not well-understood in plasma pyrolysis. This project aims to leverage advances in the kinetic modeling of soot formation in combustion, as well as in surrogate modeling and active learning, to systematically investigate the effects of process parameter on the kinetics of PAH formation in plasma pyrolysis of methane. To this end, we propose to use the PAH formation kinetics model developed by the PPPL/PU group based on the well-established ABF and HACA mechanisms, coupled with low-temperature plasma models. We will develop an active learning (AL) framework based on Bayesian optimization to systematically and data-efficiently explore the complex and multivariable parameter space of plasma pyrolysis in order to quantify the effects of plasma and feed parameters on the ABF and HACA kinetic pathways. AL is the branch of machine learning concerned with systematically querying samples from a system (experimental or computational) to train a data-driven model that maps design parameters to a performance criterion. We will use the data generated via AL to perform global sensitivity analysis, combined with uncertainty quantification, to elucidate the impact of different reaction pathways on minimizing formation of soot precursors. This study will result in an improved understanding of kinetics of PAH formation in plasma pyrolysis and can pave the way for more advanced mechanistic studies (e.g., soot nucleation mechanisms). Additionally, the findings will be useful for establishing practical strategies for increasing the pyrolysis efficiency and producing high-grade carbon for synthesis of nanomaterials.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Hybrid Data‐Driven Discovery of High‐Performance Silver Selenide‐Based Thermoelectric Composites

Optimizing material compositions often enhances thermoelectric performances. However, the large selection of possible base elements and dopants results in a vast composition design space that is too large to systematically search using solely domain knowledge. To address this challenge, a hybrid data-driven strategy that integrates Bayesian optimization (BO) and Gaussian process regression (GPR) is proposed to optimize the composition of five elements (Ag, Se, S, Cu, and Te) in AgSe-based thermoelectric materials. Data is collected from the literature to provide prior knowledge for the initial GPR model, which is updated by actively collected experimental data during the iteration between BO and experiments. Within seven iterations, the optimized AgSe-based materials prepared using a simple high-throughput ink mixing and blade coating method deliver a high power factor of 2100 µW m −1 K −2 , which is a 75% improvement from the baseline composite (nominal composition of Ag 2 Se 1 ). In conclusion, the success of this study provides opportunities to generalize the demonstrated active machine learning technique to accelerate the development and optimization of a wide range of material systems with reduced experimental trials.

36 MATERIALS SCIENCE↗

Particle Markov Chain Monte Carlo Approach to Inference in Transient Surface Kinetics

Here, in this work, we develop a novel Bayesian approach to study the adsorption and desorption of CO onto a Pd(111) surface, a process of great importance in natural sciences. The motivation for this work comes from the recent availability of time-resolved infrared spectroscopy data and the need for model interpretability and uncertainty quantification in chemical processes. The objective is to learn the relevant parameters that characterize the process: coverage with time, rate constants, activation energies, and pre-exponential factors. Our approach consists of three main schemes: (i) a problem design and probabilistic model for the whole system, (ii) a particle Markov chain Monte Carlo sampler to learn the hidden coverages and rate constant parameters, and (iii) two Bayesian formulations to infer the activation energies and pre-exponential factors. The flexibility of the Bayesian framework allows for uncertainty quantification where possible and integration of mathematical constraints in the model to reflect the system physically. We found that our results for the activation energies and pre-exponential factor are in agreement with those reported in the experimental literature, independently, and we provide discussions on the advantages and disadvantages as well as applicability to other systems.

36 MATERIALS SCIENCE↗

Bayesian calibration of irradiated graphite property models under high temperatures

Graphite under high temperatures and irradiation is central to advanced reactors. We develop a Bayesian calibration framework for graphite property models that explicitly represents model-data mismatch via a Gaussian-process discrepancy. The approach propagates uncertainty from parameters, experimental noise, and model form, with a hierarchical variance structure to capture group and cross-group noise. Using two predictive models across five grades (IG-110, NBG-18, PCEA, NBG-17, 2114) and four properties-irradiation-induced dimension change, creep, Young’s modulus change ratio, and coefficient of thermal expansion change ratio-we obtain average predictive-error reductions of 54%, 65%, 17%, and 17% when discrepancy is included. We illustrate engineering impact with a multiphysics model of a very-high-temperature reactor prismatic reflector brick, analyzing stresses under high fluence and temperature. Accounting for model discrepancy markedly improves predictive accuracy and provides a robust basis for reliable graphite component design in advanced reactors.

36 - MATERIALS SCIENCE↗

Tailoring Molecular Space to Navigate Phase Complexity in Cs-Based Quasi-2D Perovskites via Gated-Gaussian-Driven High-Throughput Discovery

Cesium-based quasi-2D halide perovskites (HPs) offer promising functionalities and low-temperature manufacturability, suited to stable tandem photovoltaics. However, the chemical interplays between the molecular spacers and the inorganic building blocks during crystallization cause substantial phase complexities in the resulting matrices. To successfully optimize and implement the quasi-2D HP functionalities, a systematic understanding of spacer chemistry, along with the seamless navigation of the inherently discrete molecular space, is necessary. Herein, by utilizing high-throughput automated experimentation, the phase complexities in the molecular space of quasi-2D HPs are explored, thus identifying the chemical roles of the spacer cations on the synthesis and functionalities of the complex materials. Furthermore, a novel active machine learning algorithm leveraging a two-stage decision-making process, called gated Gaussian process Bayesian optimization is introduced, to navigate the discrete ternary chemical space defined with two distinctive spacer molecules. Through simultaneous optimization of photoluminescence intensity and stability that “tailors” the chemistry in the molecular space, a ternary-compositional quasi-2D HP film realizing excellent optoelectronic functionalities is demonstrated. Finally, this work not only provides a pathway for the rational and bespoke design of complex HP materials but also sets the stage for accelerated materials discovery in other multifunctional systems.

36 MATERIALS SCIENCE↗

Bayesian Calibration of Nuclear Graphite Property Models Accounting for Model Inadequacy and Impacts on Component Performance

Nuclear-grade structural graphite is extensively utilized in the core designs of various advanced nuclear reactors. In the reactor environment, graphite is subjected to prolonged exposure to extreme conditions, including high temperatures, radiation, and potentially molten salt and oxygen. Such exposure can induce several degradation mechanisms in graphite, including nonuniform volumetric strains caused by irradiation and thermal expansion, leading to stresses that may compromise the performance of graphite components. Assessing component integrity requires accurate models of graphite's thermomechanical response. This report documents the Bayesian calibration of thermomechanical properties for nuclear-grade graphite and their application to graphite component modeling and simulation using the Grizzly code. As part of this work, uncertainty-quantified models were developed for the elastic modulus, coefficient of thermal expansion, irradiation-induced dimensional change, and irradiation-induced creep for graphite grades IG-110, NBG-18, NBG-17, PCEA, and 2114. Using a hierarchical Bayesian approach, multiple experimental data sources were combined to develop Gaussian process models for the properties. Using the Kennedy O'Hagan framework, the uncertainties due to inadequacies in the model and the inherent spread in the experimental data were quantified for three different models. These uncertainty-quantified models, with a model-form correction, were subsequently applied to a coupled-physics simulation of representative graphite components, revealing that the uncertainties have a large impact on the components' deformation.

36 - MATERIALS SCIENCE↗