Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Machine Learning Models”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

A Methodology for Simulating Supercritical CO2 Heat Transfer Experiments Using Machine Learning Models

To support the growth of supercritical carbon dioxide (sCO2) power cycles in the energy industry, this study seeks to train a machine learning model to mirror experimental data to predict new heat transfer data. To do this experimental data was amassed, one preliminary set comprised of 16 test results, and an expanded version comprised of 38 test results. With the goal of predicting experimental apparatus temperatures and pressures, several iterations of models were tested investigating the impact of model hyper-parameters, data inclusion, and data pre-processing on model performance. A total of 15 variations cumulatively of Gaussian Process Regressors, Gradient Boosting Regressors, and Multi-Layer Perceptrons were trained and validated on the preliminary set, and the best algorithm of each class was re-trained on the expanded set. These were compared based on test/train R^2 , test/train mean absolute error (MAE), and validation MAE, to identify the successfulness of these models. It was shown temperatures could be predicted within just a few degrees, showing the potential of this approach. Future research has been identified with approaches to improve pressure and temperature predictions going forward.

Grabowski, Owen↗

Source Analysis of Ozone Pollution in Liaoyuan City’s Atmosphere Based on Machine Learning Models and HYSPLIT Clustering Method

Firstly, this study investigates the spatiotemporal distribution characteristics of the ozone (O 3 ) pollution in Liaoyuan City using monitoring data from 2015 to 2024. Then, three machine learning models (ML)—random forest (RF), support vector machine (SVM), and artificial neural network (ANN)—are employed to quantify the influence of meteorological and non-meteorological factors on O 3 concentrations. Finally, the HYSPLIT clustering method and CMAQ model are utilized to analyze inter-regional transport characteristics, identifying the causes of O 3 pollution. The results indicate that O 3 pollution in Liaoyuan exhibits a distinct seasonal pattern, with the highest concentrations found in spring and summer, peaking in the afternoon. Among the three ML models, the random forest model demonstrates the best predictive performance (R 2 = 0.9043). Feature importance identifies NO 2 as the primary driving factor, followed by meteorological conditions in the second quarter and land surface characteristics. Furthermore, regional transport significantly contributes to O 3 pollution, with approximately 80% of air mass trajectories in heavily polluted episodes originating from adjacent industrial areas and the sea. The combined effects of transboundary precursors and O 3 transport with local emissions and meteorological conditions further increase the O 3 pollution level. This study highlights the need to strengthen coordinated NO X and VOCs emission reductions and enhance regional joint prevention and control strategies in China.

HYSPLIT clustering↗

Toward fully automated UED operation using two-stage machine learning model

To demonstrate the feasibility of automating UED operation and diagnosing the machine performance in real time, a two-stage machine learning (ML) model based on self-consistent start-to-end simulations has been implemented. This model will not only provide the machine parameters with adequate precision, toward the full automation of the UED instrument, but also make real-time electron beam information available as single-shot nondestructive diagnostics. Furthermore, based on a deep understanding of the root connection between the electron beam properties and the features of Bragg-diffraction patterns, we have applied the hidden symmetry as model constraints, successfully improving the accuracy of energy spread prediction by a factor of five and making the beam divergence prediction two times faster. The capability enabled by the global optimization via ML provides us with better opportunities for discoveries using near-parallel, bright, and ultrafast electron beams for single-shot imaging. It also enables directly visualizing the dynamics of defects and nanostructured materials, which is impossible using present electron-beam technologies.

36 MATERIALS SCIENCE↗

A Methodology for Simulating Supercritical CO2 Heat Transfer Experiments Using Machine Learning Models

In an effort to support the growth of supercritical carbon dioxide (sCO2) power cycles in the energy industry, this study seeks to train a machine learning model to mirror experimental data to inform future efforts and design features for both sCO2 heat exchangers and sCO2 turbine thermal management. There is a need for large amounts of experimental testing as there is less established literature about sCO2 used as a working medium in these cycles, as well as due to the influx of novel heat transfer designs presented by the advent of additive manufacturing.

Grabowski, Owen↗

Learning to Branch with Interpretable Machine Learning Models

This presentation describes an algorithm for applying machine learning to branching to speed up the solution of integer optimization problems. These problems are challenging and solved multiple times a day by power systems operators. We show that our approach speeds up a widely used open-source optimization solver.

Bayramoglu, Selin↗

Multiscale and Machine Learning Modeling for Process-informed Microstructure Prediction in Additively Manufactured Materials Using MALAMUTE

Advanced Materials and Manufacturing Technologies (AMMT) program under the Department of Energy Office of Nuclear Energy, aims to develop and qualify additively-manufactured materials for nuclear applications. The key challenges to these efforts are the microstructural variabilities observed on the AM products and their impact on the properties and performance of the material in extreme environments. AMMT is using a combination of high-through-put experimental and modeling techniques to accelerate the qualification efforts. Conventionally, in-situ and ex-situ characterizations and testing are performed to correlate different aspects of the AM process to the final product and its performance. However, adopting a trial-and-error approach to experimentally evaluate the vast range of process parameters required to capture the microstructural variabilities is cost-prohibitive. Modeling and simulation provide a comparatively inexpensive way to understand and correlate the microstructural evolution to the processing conditions. The modeling and simulation work-packages within the AMMT program aims to use physics-based and machine learning modeling capabilities to develop a digital twin for AM that can correlate the process conditions to the final product and establish a process-structure-property-performance (PSPP) correlation for AM materials. The melting and subsequent solidification that occurs during the AM process is a complex phenomenon that requires multiscale multiphysics analysis. Idaho National Laboratory’s (INL) Multiphysics Object-Oriented Simulation Environment (MOOSE), specifically the MOOSE Application Library for Advanced Manufacturing UTilitiEs (MALAMUTE) software, provides an ideal platform for developing the multiphysics multiscale model to explore the intricacies of the microstructural evolution during the AM processes within a single framework. Furthermore, given that such full-fidelity simulations can be computationally intensive, reduced order models are necessary to explore the PSPP space for AM materials in an efficient, reliable, and cost-effective way. This work package focuses on understanding the role of process variabilities on the various microstructural characteristics of the AM materials. Microstructures unique to AM materials, such as compositional micro-heterogeneity and dislocation cells, are of particular interest here since they can influence the creep properties and radiation performance. In fiscal year (FY) 24, we significantly advanced upon our work in the last fiscal year, both on physics-based and ML models. The alloy solidification model available in MOOSE has been extended to incorporate the thermodynamic properties and free energy relevant to 316SS. The model demonstrates the Cr segregation that occurs during solidifcation. It is demonstrated that rate of solidification and solute segregation is primarily influence by the cooling rate dictating the level of freezing. This work captures the microstructural variabilities at the subgrain level that are often missing in the part-scale models. With an aim to connect the microstructural evolution model to realistic process conditions, a reduced order model is developed for predicting the thermal conditions around meltpool from high-fidelity process simulations. Furthermore, machine learning approach is used to accelerate the temperature prediction during the AM process. In the following years, MALAMUTE will be used to connect different aspects of the models and quantitatively predict the microstructural evolution. The developed ML-based surrogate model will consider the process conditions as the input to predict the microstructural features in a cost-effective way. The generated microstructures can be used by other work packages under AMMT to evaluate the properties and environmental response of the material at the mesoscale. Thus, this work help identify the key microstructural features at the subgrain level that are significant in property/performance prediction of the AM products. This work will provide inputs to the large-scale process variability models to reevaluate and validate assumptions/simplifications made in the part-scale models. Furthermore, through active learning this work will help identify the data need from both modeling and experimental sides for development of a robust digital twin for AM.

36 MATERIALS SCIENCE↗

Chemical composition based machine learning model to predict defect formation in additive manufacturing

With a goal of exploiting additive manufacturing to improve the manufacturing of existing reactor materials, we developed a chemical composition-based machine learning model to predict the printability of any given alloy in laser powder bed fusion (L-PBF) using experimental data from peer-reviewed literature. We defined printability as the ability to avoid defects like cracking, balling, porosity, and lack of fusion, that are caused by thermal stresses (during solidification or liquation), molten pool disintegration into disconnected small beads or lack of heat input respectively. Our models predict the tendency of balling defect formation and porosity percentage for a given composition, under a given set of processing conditions. To predict the likelihood of balling defect, three models: a random forest classifier, a gradient boost regressor and a neural network were trained on a dataset containing both traditional alloys and high entropy alloys. The neural network model showed the highest accuracy of 92.3 % in predicting the balling defect formation. A random forest regressor, gradient boost regressor and neural network were trained and tested on a dataset of various alloys to predict porosity. The random forest regressor showed the best predictions with an R 2 score of 0.97. The models also revealed the relative importance of the input descriptors on defect-formation tendency. Of particular significance was the identification of carbon as an important element in determining the occurrence of balling and percent porosity in alloys like steel, as well as being moderately important to the percentage porosity in other alloys as well as steel. Manganese was also identified as a key descriptor for the percentage of porosity in steel and other alloys. Manganese’s low thermal conductivity and consistent presence in the dataset is the likely cause for its contribution. Carbon’s role is attributable to its relatively high specific heat and high melting temperature. In conclusion, our model serves as a swift, chemistry-based tool to design experiments and find modified compositions better suited for additive manufacturing.

36 MATERIALS SCIENCE↗

SNM Radiation Signature Classification Using Different Semi-Supervised Machine Learning Models

The timely detection of special nuclear material (SNM) transfers between nuclear facilities is an important monitoring objective in nuclear nonproliferation. Persistent monitoring enabled by successful detection and characterization of radiological material movements could greatly enhance the nuclear nonproliferation mission in a range of applications. Supervised machine learning can be used to signal detections when material is present if a model is trained on sufficient volumes of labeled measurements. However, the nuclear monitoring data needed to train robust machine learning models can be costly to label since radiation spectra may require strict scrutiny for characterization. Therefore, this work investigates the application of semi-supervised learning to utilize both labeled and unlabeled data. As a demonstration experiment, radiation measurements from sodium iodide (NaI) detectors are provided by the Multi-Informatics for Nuclear Operating Scenarios (MINOS) venture at Oak Ridge National Laboratory (ORNL) as sample data. Anomalous measurements are identified using a method of statistical hypothesis testing. After background estimation, an energy-dependent spectroscopic analysis is used to characterize an anomaly based on its radiation signatures. In the absence of ground-truth information, a labeling heuristic provides data necessary for training and testing machine learning models. Supervised logistic regression serves as a baseline to compare three semi-supervised machine learning models: co-training, label propagation, and a convolutional neural network (CNN). In each case, the semi-supervised models outperform logistic regression, suggesting that unlabeled data can be valuable when training and demonstrating value in semi-supervised nonproliferation implementations.

98 NUCLEAR DISARMAMENT, SAFEGUARDS, AND PHYSICAL P↗

A Comprehensive Machine Learning Model for Metal–Ligand Binding Prediction: Applications in Chemistry and Biology

A machine-learning (ML) model that predicts metal–ligand binding constants was developed using the open-source Chemprop software. The model was trained on over 30,000 experimental log K 1 values, which include both protonation and metal–ligand stability constants, comprising over 3500 ligands and 10 2 metal ions from 73 total elements, thus generalizing beyond existing limited approaches, which focus only on specific metals or ligand families. The best-performing model included a combination of SMILES-based molecular representations along with descriptors for the metal ion and experimental conditions. It had an external test R 2 value of 0.942, and MAE value of 0.834. A “SMILES-only” simpler version also produced accurate predictions and preserved the binding trends, serving as a quick and easily accessible alternative for users without computational expertise. The SMILES-only model performed comparably to density functional theory (DFT) calculations but utilized a fraction of the computational resources. The model was successfully applied across diverse domains, including bioinorganic chemistry, heavy metal remediation, and sensor development and demonstrated its effectiveness as a rapid and reliable screening tool for both academic and industrial uses.

Ligands↗

Observations and Machine-Learned Models of Near-Surface Permafrost along the Koyukuk River, Alaska, USA

This dataset contains GeoTIFs (raster) and GeoPackages (vector) that map observations of near-surface permafrost and not-permafrost from a field campaign conducted near the village of Huslia, AK along the Koyukuk River and its floodplain in July 2018. These data were collected as part of a campaign to understand if and how permafrost impacts riverbank erosion. This problem cannot be assessed without knowing where permafrost exists. Permafrost was observed via frost probing (to a maximum depth of one meter), coring (to a maximum depth of two meters) and bank/bar excavations. An additional boat survey was performed wherein expert (Joel Rowland) judgment assessed the presence or absence of distinctive permafrost features (e.g., overhanging tundra mats, thermoerosional niching, ice wedges, active drainage of ice melt from soils). This dataset also contains the input features and results of two machine learning models (random forest and convolutional neural network) that extrapolate the observations to the full floodplain that may be useful for building, testing, or validating other machine-learned permafrost models. Permafrost data are provided as georasters of the same shape and geovectors (polylines/polygons) and are all projected into EPSG:32605. All data can be visualized with a GIS (QGIS, ArcGIS, etc.).

54 ENVIRONMENTAL SCIENCES↗

Interpretable Machine Learning Models for Practical Antimonate Electrocatalyst Performance

Computationally predicting the performance of catalysts under reaction conditions is a challenging task due to the complexity of catalytic surfaces and their evolution in situ, different reaction paths, and the presence of solid-liquid interfaces in the case of electrochemistry. We demonstrate here how relatively simple machine learning models can be found that enable prediction of experimentally observed onset potentials. Inputs to our model are comprised of data from the oxygen reduction reaction on non-precious transition-metal antimony oxide nanoparticulate catalysts with a combination of experimental conditions and computationally affordable bulk atomic and electronic structural descriptors from density functional theory simulations. From human-interpretable genetic programming models, we identify key experimental descriptors and key supplemental bulk electronic and atomic structural descriptors that govern trends in onset potentials for these oxides and deduce how these descriptors should be tuned to increase onset potentials. Here, we finally validate these machine learning predictions by experimentally confirming that scandium as a dopant in nickel antimony oxide leads to a desired onset potential increase. Macroscopic experimental factors are found to be crucially important descriptors to be considered for models of catalytic performance, highlighting the important role machine learning can play here even in the presence of small datasets.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

A generative machine learning model for designing metal hydrides applied to hydrogen storage

Developing new metal hydrides is a critical step toward efficient hydrogen storage in carbon-neutral energy systems. However, existing materials databases, such as the Materials Project, contain a limited number of well-characterized hydrides, which constrains the discovery of optimal candidates. This work presents a framework that integrates causal discovery with a lightweight generative machine learning model to generate novel metal hydride candidates that may not exist in current databases. Using a dataset of 450 samples (270 training, 90 validation, and 90 testing), the model generates 1000 candidates. After ranking and filtering, six previously unreported chemical formulas and crystal structures are identified, four of which are validated by density functional theory simulations and show strong potential for future experimental investigation. Overall, the proposed framework provides a scalable and time-efficient approach for expanding hydrogen storage datasets and accelerating materials discovery.

generative model↗

Energy Prediction under Changed Demand Conditions: Robust Machine Learning Models and Input Feature Combinations

Deciding on a suitable algorithm for energy demand prediction in a building is non-trivial and depends on the availability of data. In this paper we compare four machine learning models, commonly found in the literature, in terms of their generalization performance and in terms of how using different sets of input features affects accuracy. This is tested on a data set where consumption patterns differ significantly between training and evaluation because of the Covid-19 pandemic. We provide a hands-on guide and supply a Python framework for building operators to adapt and use in their applications.

Schranz, Thomas↗

Physiochemical Machine Learning Models Predict Operational Lifetimes of CH3NH3PbI3 Perovskite Solar Cells

Halide perovskites are promising photovoltaic (PV) materials with the potential to lower the cost of electricity and greatly expand the penetration of PV if they can demonstrate long-term stability under illumination in the presence of moisture and oxygen. The solar cell service lifetime as quantified by the T80 (the time required for the power conversion efficiency to drop to 80% of its starting value) is a useful metric to assess stability. The T80 for utility, commercial, or residential PV systems needs to be several decades in order to yield low-cost electricity, and thus it is not practical to directly measure the T80. It would be useful if T80 could be predicted from the initial dynamics of a solar cell’s performance, but until now no models have been developed to forecast T80. In this work, we report the development of machine learning models to predict T80 of ITO/NiOx/CH3NH3PbI3/C60/BCP/Ag solar cells operating at maximum power point under 1-sun equivalent photon flux in air at varying temperatures and relative humidities. Efficiency losses are driven by short-circuit current and fill factor, indicating that chemical decomposition of the perovskite is a major contributor to degradation. Spatial patterns evident from in situ dark field optical microscopy suggest that the electric field gradient at device edges plays a significant role in perovskite decomposition, along with photochemical reactions with O2 and H2O. Models are trained using a menu of features from three distinct categories: (i) features based on measurements of the initial rates of change of device parameters, (ii) features based on the ambient conditions during operation (temperature, & partial pressure of H2O), and (iii) features based on underlying physics and chemistry. We show that a theory-based physiochemical feature derived from a model of the chemical reaction kinetics of the rate of degradation of the CH3NH3PbI3 is particularly valuable for prediction. This physiochemical feature was selected as the first or second most dominant feature in the best performing models. With a dataset consisting of 45 accelerated degradation experiments with T80 that range over a factor of 30, the model predicts T80 with an accuracy of about 40% (|predicted T80 - observed T80| / observed T80) on samples not used in training. This hybrid ML approach should be effective when applied to other compositions, device architectures, and advanced packaging schemes.

14 SOLAR ENERGY↗

Physiochemical Machine Learning Models Predict Operational Lifetimes of CH3NH3PbI3 Perovskite Solar Cells

Halide perovskites are promising photovoltaic (PV) materials with the potential to lower the cost of electricity and greatly expand the penetration of PV if they can demonstrate long-term stability under illumination in the presence of moisture and oxygen. The solar cell service lifetime as quantified by the T80 (the time required for the power conversion efficiency to drop to 80% of its starting value) is a useful metric to assess stability. The T80 for utility, commercial, or residential PV systems needs to be several decades in order to yield low-cost electricity, and thus it is not practical to directly measure the T80. It would be useful if T80 could be predicted from the initial dynamics of a solar cell’s performance, but until now no models have been developed to forecast T80. In this work, we report the development of machine learning models to predict T80 of ITO/NiOx/CH3NH3PbI3/C60/BCP/Ag solar cells operating at maximum power point under 1-sun equivalent photon flux in air at varying temperatures and relative humidities. Efficiency losses are driven by short-circuit current and fill factor, indicating that chemical decomposition of the perovskite is a major contributor to degradation. Spatial patterns evident from in situ dark field optical microscopy suggest that the electric field gradient at device edges plays a significant role in perovskite decomposition, along with photochemical reactions with O2 and H2O. Models are trained using a menu of features from three distinct categories: (i) features based on measurements of the initial rates of change of device parameters, (ii) features based on the ambient conditions during operation (temperature, & partial pressure of H2O), and (iii) features based on underlying physics and chemistry. We show that a theory-based physiochemical feature derived from a model of the chemical reaction kinetics of the rate of degradation of the CH3NH3PbI3 is particularly valuable for prediction. This physiochemical feature was selected as the first or second most dominant feature in the best performing models. With a dataset consisting of 45 accelerated degradation experiments with T80 that range over a factor of 30, the model predicts T80 with an accuracy of about 40% (|predicted T80 - observed T80| / observed T80) on samples not used in training. This hybrid ML approach should be effective when applied to other compositions, device architectures, and advanced packaging schemes.

14 SOLAR ENERGY↗

Predicting industrial building energy consumption with statistical and machine-learning models informed by physical system parameters

The industrial sector consumes about one-third of global energy, making them a frequent target for energy use reduction. Variation in energy usage is observed with weather conditions, as space conditioning needs to change seasonally, and with production, energy-using equipment is directly tied to production rate. Previous models were based on engineering analyses of equipment and relied on site-specific details. Others consisted of single-variable regressors that did not capture all contributions to energy consumption. Further, new modeling techniques could be applied to rectify these weaknesses. Applying data from 45 different manufacturing plants obtained from industrial energy audits, a supervised machine-learning model is developed to create a general predictor for industrial building energy consumption. The model uses features of air enthalpy, solar radiation, and wind speed to predict weather-dependency; motor, steam, and compressed air system parameters to capture support equipment contributions; and operating schedule, production rate, number of employees, and floor area to determine production-dependency. Results showed that a model that used a linear regressor over a transformed feature space could outperform a support vector machine and utilize features more representative of physical systems. Using informed parameters to build a reliable predictor will more accurately characterize a manufacturing facility's energy savings opportunities.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Multiscale and Machine Learning Modeling for Process-informed Microstructure Prediction in Additively Manufactured Materials using MALAMUTE

The Advanced Materials and Manufacturing Technologies (AMMT) program under the Department of Energy Office of Nuclear Energy aims to develop and qualify additively manufactured materials for nuclear applications. One key challenge to this is the microstructural variability observed in the additively manufactured products and their impact on the properties and performance of the material in extreme environments. AMMT is using a combination of high-throughput experimental and modeling techniques to accelerate qualification. Conventionally, in-situ and ex-situ characterizations and testing are performed to correlate different aspects of the additive manufacturing process to the final product and its performance. However, adopting a trial-and-error approach to experimentally evaluate the vast range of process parameters required to capture microstructural variability is cost-prohibitive. Modeling and simulation provide a comparatively inexpensive way to understand and correlate the microstructural evolution to the processing conditions. The modeling and simulation work-packages within the AMMT program aims to use physics-based and machine learning models to develop a digital twin for additive manufacturing that can correlate the process conditions to the final product and establish a process-structure-property-performance (PSPP) correlation. The melting and subsequent solidification that occurs during the additive process is a complex phenomenon that requires multiscale multiphysics analysis. This work package focuses on understanding the role of process variabilities on the unique microstructural characteristics of additively manufactured materials. Microstructural features at the subgrain level, such as compositional micro-heterogeneity and dislocation cells, are of particular interest here since they can influence the creep properties and radiation performance. Idaho National Laboratory’s Multiphysics Object-Oriented Simulation Environment (MOOSE), specifically the MOOSE Application Library for Advanced Manufacturing UTilitiEs (MALAMUTE) software, provides an ideal platform for developing the multiphysics multiscale model to explore the intricacies of the microstructural evolution during the AM processes within a single framework. Furthermore, given that such full-fidelity simulations can be computationally intensive, reduced order models are necessary to explore the PSPP space for additively manufactured materials in an efficient, reliable, and cost-effective way. This work focuses on capturing the microstructural variabilities at the subgrain level that are often missing in the part-scale models. In fiscal year 2025, we significantly advanced upon our work in the last fiscal year, in terms of the predictive capabilities of the physics-based and ML models, by adding the capabilities to capture subgrain-level micro-segregation during solidification using phase-field model and to predict the time-dependent dynamics of the AM process through the MOGPAR model. The alloy solidification model in MOOSE incorporates the thermodynamic properties and free energy relevant to 316 stainless steel. The model demonstrates the Cr and Ni segregation that occurs during solidification, including that the rate of solidification. The microstructural evolution model is connected to the process conditions via the surrogate model developed in this work. This enables predictions of the final microstructure in conjunctions with the manufacturing process. This work supports AMMT's rapid qualification goals by laying the foundation for an efficient and cost-effective model establishing the PSPP correlation for AM. The generated microstructures and predicted micro-segregation can be used by other work packages under AMMT to evaluate the properties and environmental response of the material at the mesoscale. Thus, this work helps to identify the key microstructural features at the subgrain level that are significant in property and performance predictions of additively manufactured components. This work will also provide inputs to the large-scale process variability models to reevaluate and validate assumptions and simplifications made in the part-scale models. Furthermore, through active learning this work can help identify the data need from both modeling and experimental sides for development of a robust digital twin for additive manufacturing and accelerate the AMMT's qualification efforts.

36 - MATERIALS SCIENCE↗