Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “model interpretability”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 199 records · Page 11

Learning thermodynamic master equations for open quantum systems

The characterization of Hamiltonians and other components of open quantum dynamical systems plays a crucial role in quantum computing and other applications. Scientific machine learning techniques have been applied to this problem in a variety of ways, including by modeling with deep neural networks. However, the majority of mathematical models describing open quantum systems are linear, and the natural nonlinearities in learnable models have not been incorporated using physical principles. We present a data-driven model for open quantum systems that includes learnable, thermodynamically consistent terms. The trained model is interpretable, as it directly estimates the system Hamiltonian and linear components of coupling to the environment. We validate the model on synthetic two and three-level data, as well as experimental two-level data collected from a quantum device at Lawrence Livermore National Laboratory.

Mathematics and Computing↗

Artificial Intelligence-Assisted Daytime Video Monitoring for Bird, Insect, and Other Wildlife Interactions with Photovoltaic Solar Energy Facilities

Studying bird, insect, and other wildlife interactions with photovoltaic (PV) solar energy facilities is difficult due to limited multi-season, multi-site data. Researchers can address such data gaps by combining passive monitoring and artificial intelligence (AI). As a part of the development of AI-enabled avian–solar monitoring software, we collected over 19,000 h of daytime videos at five PV sites across three U.S. regions between 2019 and 2024. We applied a moving object detection and tracking (MODT Version 1) AI model we developed earlier to 4373 h of the footage to extract moving objects in video frames, and human reviewers interpreted the model output and identified 68,646 bird, 25,968 insect, and 169 other wildlife instances to generate the training/validation dataset. We analyzed the data by site, region, and season, considering ground cover and landscapes. Songbirds were most common, with raptors as the next most frequent group. Most notably, no bird collisions were confirmed in our observations collected from the videos. Birds most often flew over or near panels, with the highest observations in the Midwest and Northeast (approximately 30 observations per hour on average) and fewer in the desert Southwest. Other behaviors included perching, foraging, and nesting. Bird abundance peaked during breeding and migration seasons. AI-assisted video monitoring proved effective for non-invasively studying flying wildlife at solar facilities to inform ecologically mindful energy development.

avian mortality↗

A hybrid calorimetry-simulation model of mixing enthalpy for molten salt

Calorimetric determination of enthalpies of mixing (ΔH mix ) in multicomponent molten salts is often interpreted using empirical models that lack physically meaningful parameters. However, for improving pyrochemical separation of spent nuclear fuel, where lanthanides are major fission products and critical elements, a deeper thermodynamic understanding of the link between excess thermodynamic properties and solvation structure is critically needed. In this work, we implement a hybrid and physics-informed framework, MIVM+Calorimetry+AIMD, which integrates experimentally measured ΔH mix (via high temperature drop calorimetry) with solvation structures from ab initio molecular dynamics (AIMD). This approach is demonstrated using LaCl 3 mixed with eutectic LiCl-KCl (58 mol% – 42 mol%) at 873 K and 1133 K. MIVM-derived parameters enable extrapolation of excess Gibbs energy and La 3+ activity across compositions. In contrast, direct ΔH mix predictions from AIMD and polarizable ion model simulations deviate significantly. By incorporating experimentally benchmarked solvation structures into an interpretable thermodynamic model, the MIVM+Calorimetry+AIMD formalism achieves higher accuracy and generalizable method for studying molten salts, offering a robust path for understanding and optimizing molten salt chemistry relevant to nuclear fuel cycles and separation science.

Goncharov, Vitaliy G. [Washington State Univ., Pul↗

A Gauss’s law analysis of redox active adsorbates on semiconductor electrodes: The charging and faradaic currents are not independent

A detailed framework for modeling and interpreting the data in totality from a cyclic voltammetric measurement of adsorbed redox monolayers on semiconductor electrodes has been developed. A three-layer model consisting of the semiconductor space-charge layer, a surface layer, and an electrolyte layer is presented that articulates the interplay between electrostatic, thermodynamic, and kinetic factors in the electrochemistry of a redox adsorbate on a semiconductor. Expressions are derived that describe the charging and faradaic current densities individually, and an algorithm is demonstrated that allows for the calculation of the total current density in a cyclic voltammetry measurement as a function of changes in the physical properties of the system (e.g., surface recombination, dielectric property of the surface layer, and electrolyte concentration). The most profound point from this analysis is that the faradaic and charging current densities can be coupled. That is, the common assumption that these contributions to the total current are always independent is not accurate. Their interrelation can influence the interpretation of the charge-transfer kinetics under certain experimental conditions. More generally, this work not only fills a long-standing knowledge gap in electrochemistry but also aids practitioners advancing energy conversion/storage strategies based on redox adsorbates on semiconductor electrodes.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Differentiable modelling to unify machine learning and physical models for geosciences

Process-based modelling offers interpretability and physical consistency in many domains of geosciences but struggles to leverage large datasets efficiently. Machine-learning methods, especially deep networks, have strong predictive skills yet are unable to answer specific scientific questions. Here, in this Perspective, we explore differentiable modelling as a pathway to dissolve the perceived barrier between process-based modelling and machine learning in the geosciences and demonstrate its potential with examples from hydrological modelling. ‘Differentiable’ refers to accurately and efficiently calculating gradients with respect to model variables or parameters, enabling the discovery of high-dimensional unknown relationships. Differentiable modelling involves connecting (flexible amounts of) prior physical knowledge to neural networks, pushing the boundary of physics-informed machine learning. It offers better interpretability, generalizability, and extrapolation capabilities than purely data-driven machine learning, achieving a similar level of accuracy while requiring less training data. Additionally, the performance and efficiency of differentiable models scale well with increasing data volumes. Under data-scarce scenarios, differentiable models have outperformed machine-learning models in producing short-term dynamics and decadal-scale trends owing to the imposed physical constraints. Differentiable modelling approaches are primed to enable geoscientists to ask questions, test hypotheses, and discover unrecognized physical relationships. Future work should address computational challenges, reduce uncertainty, and verify the physical significance of outputs.

58 GEOSCIENCES↗

A search for new resonances in multiple final states with a high transverse momentum Z boson in $\sqrt{s} $ = 13 TeV pp collisions with the ATLAS detector

A generic search for resonances is performed with events containing a Z boson with transverse momentum greater than 100 GeV, decaying into e + e – or μ + μ – . The analysed data collected with the ATLAS detector in proton-proton collisions at a centre-of-mass energy of 13 TeV at the Large Hadron Collider correspond to an integrated luminosity of 139 fb –1 . Two invariant mass distributions are examined for a localised excess relative to the expected Standard Model background in six independent event categories (and their inclusive sum) to increase the sensitivity. No significant excess is observed. Exclusion limits at 95% confidence level are derived for two cases: a model-independent interpretation of Gaussian-shaped resonances with the mass width between 3% and 10% of the resonance mass, and a specific heavy vector triplet model with the decay mode W' → ZW → ℓℓqq.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

maslov-group/FUN-PROSE

mRNA levels of all genes in a genome is a critical piece of information defining the overall state of the cell in a given environmental condition. Being able to reconstruct such condition-specific expression in fungal genomes is particularly important to metabolically engineer these organisms to produce desired chemicals in industrially scalable conditions. Most previous deep learning approaches focused on predicting the average expression levels of a gene based on its promoter sequence, ignoring its variation across different conditions. Here we present FUN-PROSE—a deep learning model trained to predict differential expression of individual genes across various conditions using their promoter sequences and expression levels of all transcription factors. We train and test our model on three fungal species and get the correlation between predicted and observed condition-specific gene expression as high as 0.85. We then interpret our model to extract promoter sequence motifs responsible for variable expression of individual genes. We also carried out input feature importance analysis to connect individual transcription factors to their gene targets. A sizeable fraction of both sequence motifs and TF-gene interactions learned by our model agree with previously known biological information, while the rest corresponds to either novel biological facts or indirect correlations.

Liu, Simon↗

FUN-PROSE: A deep learning approach to predict condition-specific gene expression in fungi

mRNA levels of all genes in a genome is a critical piece of information defining the overall state of the cell in a given environmental condition. Being able to reconstruct such condition-specific expression in fungal genomes is particularly important to metabolically engineer these organisms to produce desired chemicals in industrially scalable conditions. Most previous deep learning approaches focused on predicting the average expression levels of a gene based on its promoter sequence, ignoring its variation across different conditions. Here we present FUN-PROSE—a deep learning model trained to predict differential expression of individual genes across various conditions using their promoter sequences and expression levels of all transcription factors. We train and test our model on three fungal species and get the correlation between predicted and observed condition-specific gene expression as high as 0.85. We then interpret our model to extract promoter sequence motifs responsible for variable expression of individual genes. We also carried out input feature importance analysis to connect individual transcription factors to their gene targets. A sizeable fraction of both sequence motifs and TF-gene interactions learned by our model agree with previously known biological information, while the rest corresponds to either novel biological facts or indirect correlations.

59 BASIC BIOLOGICAL SCIENCES↗

Using Machine Learning to Understand Electric and Hybrid Vehicles Ownership in Burdened and Nonburdened Communities

Transitioning to electric and hybrid vehicles (EHVs) for all communities is a pivotal step toward sustainable transportation and environmental conservation. This paper aims to understand the adoption of EHVs, focusing on burdened communities (BCs) in the United States. The EHV ownership-based analysis combines two datasets—behavioral data from the Puget Sound Regional Travel Survey integrated with BCs (Justice40) data covering transportation insecurity, environmental burden, social vulnerability, health vulnerability, and climate and disaster risk burden. After creating this unique database, descriptive analysis and modeling are used to analyze the data and predict EHV ownership in the future. Specifically, we use a new method that combines particle swarm optimization (PSO) with a stacking model named PSO-Stacking, which incorporates heterogeneous base learners of machine learning and deep learning. PSO applies a customized objective function to select the optimal hyperparameters for heterogeneous learners within the stacking model, effectively addressing challenges such as multicollinearity, data imbalance, nonlinearity, and overfitting. The proposed solution covers more accurate results than standard benchmark models for EHV ownership in BCs and non-BCs. In addition, the results of the PSO-Stacking method are explained using the local interpretable model-agnostic explanations technique. Results show a negative correlation between the BCs indicators, that is, higher transportation insecurity associated with lower EHV ownership. Furthermore, BCs have higher future climate risk scores, diesel particulate matter levels, and PM2.5 in the air than non-BCs because of higher conventional vehicle ownership. These communities are at higher risk and can benefit from electrification, EV infrastructure, and EV policies to address environmental challenges.

Aslam, Zeeshan [ORNL]↗

A comparative study of multimodal data fusion strategies for planetary spectroscopy

Integrating heterogeneous data sources can improve scientific inference when different modalities capture complementary information, but doing so is challenging in high-dimensional, small-sample settings. In spectroscopy for planetary exploration, Laser-Induced Breakdown Spectroscopy (LIBS), Raman Spectroscopy (Raman), Visible Infrared Spectroscopy (VISIR), and Mid-Infrared Spectroscopy (MIR) each examine different aspects of composition and mineralogy, raising fundamental questions about when and how data fusion improves predictive performance. Using a Mars-relevant set of geologic standards with measurements from all four modalities, we present a rigorous systematic evaluation of four data fusion strategies: low-level (data) fusion, mid-level (feature) fusion, high-level (decision) fusion, and residual-boosting (sequential) fusion. We assess performance in predicting oxide composition via nested cross-validation and corrected significance testing to evaluate whether data fusion improves upon single-modality baselines. We show that data fusion does not uniformly improve accuracy, and that observed gains are modest, oxide-dependent, and sensitive to modality and model structure. To move beyond aggregate accuracy metrics, we use model coefficients, permutation importance, and residual gain analysis to examine how the fusion models weight individual modalities and to identify patterns of apparent complementarity or redundancy. Though focused on spectroscopy for planetary exploration, our framework for data fusion evaluation and interpretation extends to other scientific domains with heterogeneous and scarce data and provides a principled approach evaluating data fusion strategies, interpreting modality contributions, and understanding tradeoffs among data fusion strategies.

97 MATHEMATICS AND COMPUTING↗

Development, characterization, and modeling of a high-performance Ru/B2CA catalyst for ammonia synthesis

This paper documents the development and performance of a nano-phase Ru catalyst on a (BaO) x (CaO) y (Al 2 O 3 ) support. Extensive screening of the support’s ternary composition shows the best stoichiometry is (BaO) 2 (CaO)(Al 2 O 3 ), denoted B2CA. The paper first describes catalyst preparation and characterization. The paper reports a detailed 12-step reaction mechanism that represents ammonia synthesis over wide ranges of temperature, pressure, space velocity, and feed composition. Additionally, the mechanism is developed and validated using results of packed-bed experiments. The elementary reaction pathways consider surface adsorbates, including catalyst-poisoning behaviors. The rate expressions include important coverage-dependent activation barriers. Machine learning models assist interpretation of the catalyst-support interactions. The detailed chemistry is much more predictive than is possible with global representations (N 2 + 3H 2 ⇌ 2NH 3 ). The validated models can be applied to assist optimizing reactor design and operating conditions.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Digitalization of an experimental electrochemical reactor via the smart manufacturing innovation platform

The exponential increase in data produced over the last two decades has revolutionized the way we collect, store, process, analyze, model, and interpret information to improve profitability. Manufacturing is no exception. How- ever, Smart Manufacturing, the digital practice, organization, workforce, and infrastructure transformation for collection and deployment of data and models at scale and at all levels of manufacturing, is a complex, costly, and labor-intensive journey that is still seeing slow adoption. The Clean Energy Smart Manufacturing Innovation Institute (CESMII), a national Manufacturing USA public-private partnership sponsored by the Department of Energy, is addressing this scaled use of data and modeling in manufacturing. CESMII has focused on how to col- lect and use operating data for numerous applications that improve productivity, precision, and performance of manufacturing operations from factory floor to supply chain using process simulation, predictive analytics, mon- itoring and control, and real-time optimization. Because contextualized data are key, CESMII has developed the Smart Manufacturing Innovation Platform (SMIP) to lower the barriers to the data that are needed to accelerate data-based model building, improve data visualization, and more quickly gain insights. Reusable, standards-based ways of doing data collection, ingestion, and contextualization are particularly important for scaling access and use of data. The SMIP uses a standards-based definition and construct for reusable information models called an SM Profile. When an SM Profile is used in conjunction with the SMIP, the SMIP ensures the availability of contextualized, operational data for model building. The present work demonstrates Smart Manufacturing and the application of the SMIP for building several data-centered models for the operation and control of an ex- perimental electrochemical reactor that reduces carbon dioxide (CO 2 ) gas to valuable liquid and gas chemicals, such as alcohols, olefins, and syngas. We describe how the SMIP plays a central role in more effective model building and we demonstrate how the electochemical reactor can be controlled and optimized for the desired products. Use of the SMIP involves the transmission of real-time sensor measurements to a cloud resource so that the operating data are available to all model building experts. The data collection and transmission process is fully automated to greatly reduce the need for manual manipulation of the data. Data-driven machine learning models are used for advanced real-time state estimation, real-time optimization, and model-based feedback control for the reactor. The application models are implemented as a system to monitor the data flow and control the electrochemical reactor with a single visualization interface. SM Profiles are used to demonstrate reusability of the information models for the reactor and the instrumentation. The application packages, algorithms, and user interfaces developed are cast as Docker images in a library to facilitate reusability of the application models.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Revising the dynamic energy budget theory with a new reserve mobilization rule and three example applications to bacterial growth

Dynamic energy budget (DEB) theory has been applied to model a wide range of organisms, including microbes. In the standard DEB model, biomass is partitioned into reserve and structural compartments, where reserve biomass is mobilized in a pseudolinear manner (while the reserve biomass density, defined as the ratio between reserve and structural biomass, decays linearly) to drive maintenance and the growth of structural biomass (and extracellular enzyme production if it is considered). However, the linear dynamics of the reserve biomass density makes the standard DEB model incapable of explaining the slowdown of microbial growth at high reserve density that is caused by macromolecular crowding effect which reduces biochemical reaction rates (a typical situation occurs when microbes are experiencing severe moisture stress) and is inconsistent with the observation that intracellular enzymatic reactions generally follow non-linear kinetics. By partitioning biomass into reserve, kinetic, and structural compartments, we show here that the Equilibrium Chemistry Approximation (ECA) kinetics can be used to represent enzymatically catalyzed reserve biomass mobilization that can then drive the kinetic and structural biomass synthesis. This revised DEB model better represents the tradeoff in ribosome allocation for structural growth and internal enzyme production, is structurally compatible with metabolic models of cell individuals, and includes the standard DEB model and the popular compromise model as special cases for representing population growth. We then applied the revised DEB model to interpret components of bacterial respiration, their dependence on substrate availability, and emergent microbial carbon use efficiency dynamics for an exponentially growing population. We found that the revised DEB model enables a better understanding of bacterial substrates use (carbon in our examples) than that can be derived from a few other models in the literature. In particular, the revised DEB model explains why carbon use efficiency may first increase, then plateau, and finally decrease with growth rate (and substrate uptake rate), as a function of proteomics. Additionally, the revised DEB model explains why the kinetic biomass compartment needs to be divided to reasonably incorporate proteomic control of microbial growth.

59 BASIC BIOLOGICAL SCIENCES↗

Di-CNN: Domain-Knowledge-Informed Convolutional Neural Network for Manufacturing Quality Prediction

In manufacturing, convolutional neural networks (CNNs) are widely used on image sensor data for data-driven process monitoring and quality prediction. However, as purely data-driven models, CNNs do not integrate physical measures or practical considerations into the model structure or training procedure. Consequently, CNNs’ prediction accuracy can be limited, and model outputs may be hard to interpret practically. This study aims to leverage manufacturing domain knowledge to improve the accuracy and interpretability of CNNs in quality prediction. A novel CNN model, named Di-CNN, was developed that learns from both design-stage information (such as working condition and operational mode) and real-time sensor data, and adaptively weighs these data sources during model training. It exploits domain knowledge to guide model training, thus improving prediction accuracy and model interpretability. A case study on resistance spot welding, a popular lightweight metal-joining process for automotive manufacturing, compared the performance of (1) a Di-CNN with adaptive weights (the proposed model), (2) a Di-CNN without adaptive weights, and (3) a conventional CNN. The quality prediction results were measured with the mean squared error (MSE) over sixfold cross-validation. Model (1) achieved a mean MSE of 6.8866 and a median MSE of 6.1916, Model (2) achieved 13.6171 and 13.1343, and Model (3) achieved 27.2935 and 25.6117, demonstrating the superior performance of the proposed model.

47 OTHER INSTRUMENTATION↗

SULI Oral Presentation

Furthering our understanding of the prevalence and severity of issues that customers face when charging their electric vehicles (EVs) is crucial in order to improve the charging experience across the United States. This project utilizes web-scraping, machine leaning (ML), and natural language processing (NLP) techniques to analyze and categorize user-generated reviews. Selenium was used to build a data collection tool that can scrape vast amounts of user review data from the PlugShare website. Sentiment analysis was employed on this dataset in order to filter out negative reviews for further analysis. NLP techniques such as tokenization and word embedding were then used to convert user-written comments into a numerical format that a ML model can interpret. Multiple ML approaches are currently being explored in order to identify and categorize the charging issues being talked about in each review. Ultimately, the results from the ML model will be visualized and explained in a report on customer pain points to be delivered to the ChargeX Consortium, therefore revealing specific areas for improvement in the customer charging experience.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

SULI Oral Presentation

Furthering our understanding of the prevalence and severity of issues that customers face when charging their electric vehicles (EVs) is crucial in order to improve the charging experience across the United States. This project utilizes web-scraping, machine leaning (ML), and natural language processing (NLP) techniques to analyze and categorize user-generated reviews. Selenium was used to build a data collection tool that can scrape vast amounts of user review data from the PlugShare website. Sentiment analysis was employed on this dataset in order to filter out negative reviews for further analysis. NLP techniques such as tokenization and word embedding were then used to convert user-written comments into a numerical format that a ML model can interpret. Multiple ML approaches are currently being explored in order to identify and categorize the charging issues being talked about in each review. Ultimately, the results from the ML model will be visualized and explained in a report on customer pain points to be delivered to the ChargeX Consortium, therefore revealing specific areas for improvement in the customer charging experience.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Data-driven high-dimensional statistical inference with generative models

Crucial to many measurements at the LHC is the use of correlated multi-dimensional information to distinguish rare processes from large backgrounds, which is complicated by the poor modeling of many of the crucial backgrounds in Monte Carlo simulations. In this work, we introduce HI-SIGMA, a method to perform unbinned high-dimensional statistical inference with data-driven background distributions. In contradistinction to many applications of Simulation Based Inference in High Energy Physics, HI-SIGMA relies on generative ML models, rather than classifiers, to learn the signal and background distributions in the high-dimensional space. These ML models allow for interpretable inference while also incorporating model errors and other sources of systematic uncertainties. We showcase this methodology on a simplified version of a di-Higgs measurement in the bbγγ final state, where the di-photon resonance allows for background interpolation from sidebands into the signal region. We demonstrate that HI-SIGMA provides improved sensitivity as compared to standard classifier-based methods, and that systematic uncertainties can be straightforwardly incorporated by extending methods which have been used for histogram based analyses.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Search for supersymmetry using vector boson fusion signatures and missing transverse momentum in pp collisions at $\sqrt{s}$ = 13 TeV with the ATLAS detector

This paper presents a search for supersymmetric particles in models with highly compressed mass spectra, in events consistent with being produced through vector boson fusion. The search uses 140 fb −1 of proton-proton collision data at $\sqrt{s}$ = 13 TeV collected by the ATLAS experiment at the Large Hadron Collider. Events containing at least two jets with a large gap in pseudorapidity, large missing transverse momentum, and no reconstructed leptons are selected. A boosted decision tree is used to separate events consistent with the production of supersymmetric particles from those due to Standard Model backgrounds. The data are found to be consistent with Standard Model predictions. The results are interpreted using simplified models of R-parity-conserving supersymmetry in which the lightest supersymmetric partner is a bino-like neutralino with a mass similar to that of the lightest chargino and second-to-lightest neutralino, both of which are wino-like. Lower limits at 95% confidence level on the masses of next-to-lightest supersymmetric partners in this simplified model are established between 117 and 120 GeV when the lightest supersymmetric partners are within 1 GeV in mass.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗