Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “model interpretability”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 343 records · Page 19

Using Machine Learning to Understand Electric and Hybrid Vehicles Ownership in Burdened and Nonburdened Communities

Transitioning to electric and hybrid vehicles (EHVs) for all communities is a pivotal step toward sustainable transportation and environmental conservation. This paper aims to understand the adoption of EHVs, focusing on burdened communities (BCs) in the United States. The EHV ownership-based analysis combines two datasets—behavioral data from the Puget Sound Regional Travel Survey integrated with BCs (Justice40) data covering transportation insecurity, environmental burden, social vulnerability, health vulnerability, and climate and disaster risk burden. After creating this unique database, descriptive analysis and modeling are used to analyze the data and predict EHV ownership in the future. Specifically, we use a new method that combines particle swarm optimization (PSO) with a stacking model named PSO-Stacking, which incorporates heterogeneous base learners of machine learning and deep learning. PSO applies a customized objective function to select the optimal hyperparameters for heterogeneous learners within the stacking model, effectively addressing challenges such as multicollinearity, data imbalance, nonlinearity, and overfitting. The proposed solution covers more accurate results than standard benchmark models for EHV ownership in BCs and non-BCs. In addition, the results of the PSO-Stacking method are explained using the local interpretable model-agnostic explanations technique. Results show a negative correlation between the BCs indicators, that is, higher transportation insecurity associated with lower EHV ownership. Furthermore, BCs have higher future climate risk scores, diesel particulate matter levels, and PM2.5 in the air than non-BCs because of higher conventional vehicle ownership. These communities are at higher risk and can benefit from electrification, EV infrastructure, and EV policies to address environmental challenges.

Aslam, Zeeshan [ORNL]↗

Idealized model for the radial gradient of modulated cosmic rays.

Development of a conceptually idealized model for interpreting existing and future measurements of the radial gradient of modulated cosmic rays in the limit that deceleration and convection effects become comparable. The proposed model assumes that the modulated spectrum consists of two idealized regions: at low energy, below some critical transition energy T sub c, a region where the shape of the spectrum is assumed to be determined wholly by a balance between convection and deceleration effects; and, at higher energies, above T sub c, a region where the magnitude and shape of the spectrum are taken to be those given by the usual convection-diffusion modulation. The application of this model, invoking some straightforward assumptions about the behavior of T sub c, suggests that the radial gradient in the regime where deceleration processes are dominant may be related to the spectral index of the unmodulated spectrum at the energy T sub c and to the absolute level of modulation. This is in contrast to the gradient at high energies, above T sub c, which is related only to local interplanetary characteristics through convection and diffusion.

O'Gallagher, J. J.↗

Dynamic models of flux tubes in the interpretation of polarization measurements

Recent observations of Stokes parameter profiles indicate the presence of mass motions with large velocity gradients associated with small-scale magnetic elements. Dynamic models of flux tubes were used in order to interpret observations of unresolved elements. It is clear that the physical picture of the dynamic models will be quite different from the hydrostatic ones since there is a strong coupling between the magnetic and the velocity field. Polarization measurements have to be interpreted in terms of dynamic models. Two-D steady flow solutions in slender magnetic tubes have been worked out. It was found that the main properties of the intensity line profiles as well as the asymmetries of the V Stokes profiles can be explained best in terms of magnetic elements with moderate field strength.

Ribes-Nesme, E.↗

Experimental study of the thermal-acoustic efficiency in a long turbulent diffusion-flame burner

A two-year study of noise production in a long tubular burner is described. The research was motivated by an interest in understanding and eventually reducing core noise in gas turbine engines. The general approach is to employ an acoustic source/propagation model to interpret the sound pressure spectrum in the acoustic far field of the burner in terms of the source spectrum that must have produced it. In the model the sources are assumed to be due uniquely to the unsteady component of combustion heat release; thus only direct combustion-noise is considered. The source spectrum is then the variation with frequency of the thermal-acoustic efficiency, defined as the fraction of combustion heat release which is converted into acoustic energy at a given frequency. The thrust of the research was to study the variation of the source spectrum with the design and operating parameters of the burner.

Mahan, J. R.↗

A comparative study of multimodal data fusion strategies for planetary spectroscopy

Integrating heterogeneous data sources can improve scientific inference when different modalities capture complementary information, but doing so is challenging in high-dimensional, small-sample settings. In spectroscopy for planetary exploration, Laser-Induced Breakdown Spectroscopy (LIBS), Raman Spectroscopy (Raman), Visible Infrared Spectroscopy (VISIR), and Mid-Infrared Spectroscopy (MIR) each examine different aspects of composition and mineralogy, raising fundamental questions about when and how data fusion improves predictive performance. Using a Mars-relevant set of geologic standards with measurements from all four modalities, we present a rigorous systematic evaluation of four data fusion strategies: low-level (data) fusion, mid-level (feature) fusion, high-level (decision) fusion, and residual-boosting (sequential) fusion. We assess performance in predicting oxide composition via nested cross-validation and corrected significance testing to evaluate whether data fusion improves upon single-modality baselines. We show that data fusion does not uniformly improve accuracy, and that observed gains are modest, oxide-dependent, and sensitive to modality and model structure. To move beyond aggregate accuracy metrics, we use model coefficients, permutation importance, and residual gain analysis to examine how the fusion models weight individual modalities and to identify patterns of apparent complementarity or redundancy. Though focused on spectroscopy for planetary exploration, our framework for data fusion evaluation and interpretation extends to other scientific domains with heterogeneous and scarce data and provides a principled approach evaluating data fusion strategies, interpreting modality contributions, and understanding tradeoffs among data fusion strategies.

97 MATHEMATICS AND COMPUTING↗

Development, characterization, and modeling of a high-performance Ru/B2CA catalyst for ammonia synthesis

This paper documents the development and performance of a nano-phase Ru catalyst on a (BaO) x (CaO) y (Al 2 O 3 ) support. Extensive screening of the support’s ternary composition shows the best stoichiometry is (BaO) 2 (CaO)(Al 2 O 3 ), denoted B2CA. The paper first describes catalyst preparation and characterization. The paper reports a detailed 12-step reaction mechanism that represents ammonia synthesis over wide ranges of temperature, pressure, space velocity, and feed composition. Additionally, the mechanism is developed and validated using results of packed-bed experiments. The elementary reaction pathways consider surface adsorbates, including catalyst-poisoning behaviors. The rate expressions include important coverage-dependent activation barriers. Machine learning models assist interpretation of the catalyst-support interactions. The detailed chemistry is much more predictive than is possible with global representations (N 2 + 3H 2 ⇌ 2NH 3 ). The validated models can be applied to assist optimizing reactor design and operating conditions.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Digitalization of an experimental electrochemical reactor via the smart manufacturing innovation platform

The exponential increase in data produced over the last two decades has revolutionized the way we collect, store, process, analyze, model, and interpret information to improve profitability. Manufacturing is no exception. How- ever, Smart Manufacturing, the digital practice, organization, workforce, and infrastructure transformation for collection and deployment of data and models at scale and at all levels of manufacturing, is a complex, costly, and labor-intensive journey that is still seeing slow adoption. The Clean Energy Smart Manufacturing Innovation Institute (CESMII), a national Manufacturing USA public-private partnership sponsored by the Department of Energy, is addressing this scaled use of data and modeling in manufacturing. CESMII has focused on how to col- lect and use operating data for numerous applications that improve productivity, precision, and performance of manufacturing operations from factory floor to supply chain using process simulation, predictive analytics, mon- itoring and control, and real-time optimization. Because contextualized data are key, CESMII has developed the Smart Manufacturing Innovation Platform (SMIP) to lower the barriers to the data that are needed to accelerate data-based model building, improve data visualization, and more quickly gain insights. Reusable, standards-based ways of doing data collection, ingestion, and contextualization are particularly important for scaling access and use of data. The SMIP uses a standards-based definition and construct for reusable information models called an SM Profile. When an SM Profile is used in conjunction with the SMIP, the SMIP ensures the availability of contextualized, operational data for model building. The present work demonstrates Smart Manufacturing and the application of the SMIP for building several data-centered models for the operation and control of an ex- perimental electrochemical reactor that reduces carbon dioxide (CO 2 ) gas to valuable liquid and gas chemicals, such as alcohols, olefins, and syngas. We describe how the SMIP plays a central role in more effective model building and we demonstrate how the electochemical reactor can be controlled and optimized for the desired products. Use of the SMIP involves the transmission of real-time sensor measurements to a cloud resource so that the operating data are available to all model building experts. The data collection and transmission process is fully automated to greatly reduce the need for manual manipulation of the data. Data-driven machine learning models are used for advanced real-time state estimation, real-time optimization, and model-based feedback control for the reactor. The application models are implemented as a system to monitor the data flow and control the electrochemical reactor with a single visualization interface. SM Profiles are used to demonstrate reusability of the information models for the reactor and the instrumentation. The application packages, algorithms, and user interfaces developed are cast as Docker images in a library to facilitate reusability of the application models.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Revising the dynamic energy budget theory with a new reserve mobilization rule and three example applications to bacterial growth

Dynamic energy budget (DEB) theory has been applied to model a wide range of organisms, including microbes. In the standard DEB model, biomass is partitioned into reserve and structural compartments, where reserve biomass is mobilized in a pseudolinear manner (while the reserve biomass density, defined as the ratio between reserve and structural biomass, decays linearly) to drive maintenance and the growth of structural biomass (and extracellular enzyme production if it is considered). However, the linear dynamics of the reserve biomass density makes the standard DEB model incapable of explaining the slowdown of microbial growth at high reserve density that is caused by macromolecular crowding effect which reduces biochemical reaction rates (a typical situation occurs when microbes are experiencing severe moisture stress) and is inconsistent with the observation that intracellular enzymatic reactions generally follow non-linear kinetics. By partitioning biomass into reserve, kinetic, and structural compartments, we show here that the Equilibrium Chemistry Approximation (ECA) kinetics can be used to represent enzymatically catalyzed reserve biomass mobilization that can then drive the kinetic and structural biomass synthesis. This revised DEB model better represents the tradeoff in ribosome allocation for structural growth and internal enzyme production, is structurally compatible with metabolic models of cell individuals, and includes the standard DEB model and the popular compromise model as special cases for representing population growth. We then applied the revised DEB model to interpret components of bacterial respiration, their dependence on substrate availability, and emergent microbial carbon use efficiency dynamics for an exponentially growing population. We found that the revised DEB model enables a better understanding of bacterial substrates use (carbon in our examples) than that can be derived from a few other models in the literature. In particular, the revised DEB model explains why carbon use efficiency may first increase, then plateau, and finally decrease with growth rate (and substrate uptake rate), as a function of proteomics. Additionally, the revised DEB model explains why the kinetic biomass compartment needs to be divided to reasonably incorporate proteomic control of microbial growth.

59 BASIC BIOLOGICAL SCIENCES↗

Di-CNN: Domain-Knowledge-Informed Convolutional Neural Network for Manufacturing Quality Prediction

In manufacturing, convolutional neural networks (CNNs) are widely used on image sensor data for data-driven process monitoring and quality prediction. However, as purely data-driven models, CNNs do not integrate physical measures or practical considerations into the model structure or training procedure. Consequently, CNNs’ prediction accuracy can be limited, and model outputs may be hard to interpret practically. This study aims to leverage manufacturing domain knowledge to improve the accuracy and interpretability of CNNs in quality prediction. A novel CNN model, named Di-CNN, was developed that learns from both design-stage information (such as working condition and operational mode) and real-time sensor data, and adaptively weighs these data sources during model training. It exploits domain knowledge to guide model training, thus improving prediction accuracy and model interpretability. A case study on resistance spot welding, a popular lightweight metal-joining process for automotive manufacturing, compared the performance of (1) a Di-CNN with adaptive weights (the proposed model), (2) a Di-CNN without adaptive weights, and (3) a conventional CNN. The quality prediction results were measured with the mean squared error (MSE) over sixfold cross-validation. Model (1) achieved a mean MSE of 6.8866 and a median MSE of 6.1916, Model (2) achieved 13.6171 and 13.1343, and Model (3) achieved 27.2935 and 25.6117, demonstrating the superior performance of the proposed model.

47 OTHER INSTRUMENTATION↗

Analysis of EIT/LASCO Observations Using Available MHD Models: Investigation of CME Initiation Propagation and Geoeffectiveness

The Sun's activity drives the variability of geospace (i.e., near-earth environment). Observations show that the ejection of plasma from the sun, called coronal mass ejections (CMEs), are the major cause of geomagnetic storms. This global-scale solar dynamical feature of coronal mass ejection was discovered almost three decades ago by the use of space-borne coronagraphs (OSO-7, Skylab/ATM and P78-1). Significant progress has been made in understanding the physical nature of the CMEs. Observations show that these global-scale CMEs have size in the order of a solar radius (approximately 6.7 x 10(exp 5) km) near the sun, and each event involves a mass of about 10(exp 15) g and an energy comparable to that of a large flare on the order of 10(exp 32) ergs. The radial propagation speeds of CMEs have a wide range from tens to thousands of kilometers per second. Thus, the transit time to near earth's environment [i.e., 1 AU (astronomical unit)] can be as fast as 40 hours to 100 hours. The typical transit time for geoeffective events is approximately 60-80 h. This paper consists of two parts: 1) A summary of the observed CMEs from Skylab to the present SOHO will be presented. Special attention will be made to SOHO/ LASCO/ EIT observations and their characteristics leading to a geoeffectiv a CME 2) The chronological development of theory and models to interpret the physical nature of this fascinating phenomenon will be reviewed. Finally, an example will be presented to illustrate the geoeffectiveness of the CMEs by using both observation and model.

Wu, S. T.↗

SULI Oral Presentation

Furthering our understanding of the prevalence and severity of issues that customers face when charging their electric vehicles (EVs) is crucial in order to improve the charging experience across the United States. This project utilizes web-scraping, machine leaning (ML), and natural language processing (NLP) techniques to analyze and categorize user-generated reviews. Selenium was used to build a data collection tool that can scrape vast amounts of user review data from the PlugShare website. Sentiment analysis was employed on this dataset in order to filter out negative reviews for further analysis. NLP techniques such as tokenization and word embedding were then used to convert user-written comments into a numerical format that a ML model can interpret. Multiple ML approaches are currently being explored in order to identify and categorize the charging issues being talked about in each review. Ultimately, the results from the ML model will be visualized and explained in a report on customer pain points to be delivered to the ChargeX Consortium, therefore revealing specific areas for improvement in the customer charging experience.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

SULI Oral Presentation

Furthering our understanding of the prevalence and severity of issues that customers face when charging their electric vehicles (EVs) is crucial in order to improve the charging experience across the United States. This project utilizes web-scraping, machine leaning (ML), and natural language processing (NLP) techniques to analyze and categorize user-generated reviews. Selenium was used to build a data collection tool that can scrape vast amounts of user review data from the PlugShare website. Sentiment analysis was employed on this dataset in order to filter out negative reviews for further analysis. NLP techniques such as tokenization and word embedding were then used to convert user-written comments into a numerical format that a ML model can interpret. Multiple ML approaches are currently being explored in order to identify and categorize the charging issues being talked about in each review. Ultimately, the results from the ML model will be visualized and explained in a report on customer pain points to be delivered to the ChargeX Consortium, therefore revealing specific areas for improvement in the customer charging experience.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Data-driven high-dimensional statistical inference with generative models

Crucial to many measurements at the LHC is the use of correlated multi-dimensional information to distinguish rare processes from large backgrounds, which is complicated by the poor modeling of many of the crucial backgrounds in Monte Carlo simulations. In this work, we introduce HI-SIGMA, a method to perform unbinned high-dimensional statistical inference with data-driven background distributions. In contradistinction to many applications of Simulation Based Inference in High Energy Physics, HI-SIGMA relies on generative ML models, rather than classifiers, to learn the signal and background distributions in the high-dimensional space. These ML models allow for interpretable inference while also incorporating model errors and other sources of systematic uncertainties. We showcase this methodology on a simplified version of a di-Higgs measurement in the bbγγ final state, where the di-photon resonance allows for background interpolation from sidebands into the signal region. We demonstrate that HI-SIGMA provides improved sensitivity as compared to standard classifier-based methods, and that systematic uncertainties can be straightforwardly incorporated by extending methods which have been used for histogram based analyses.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Search for supersymmetry using vector boson fusion signatures and missing transverse momentum in pp collisions at $\sqrt{s}$ = 13 TeV with the ATLAS detector

This paper presents a search for supersymmetric particles in models with highly compressed mass spectra, in events consistent with being produced through vector boson fusion. The search uses 140 fb −1 of proton-proton collision data at $\sqrt{s}$ = 13 TeV collected by the ATLAS experiment at the Large Hadron Collider. Events containing at least two jets with a large gap in pseudorapidity, large missing transverse momentum, and no reconstructed leptons are selected. A boosted decision tree is used to separate events consistent with the production of supersymmetric particles from those due to Standard Model backgrounds. The data are found to be consistent with Standard Model predictions. The results are interpreted using simplified models of R-parity-conserving supersymmetry in which the lightest supersymmetric partner is a bino-like neutralino with a mass similar to that of the lightest chargino and second-to-lightest neutralino, both of which are wino-like. Lower limits at 95% confidence level on the masses of next-to-lightest supersymmetric partners in this simplified model are established between 117 and 120 GeV when the lightest supersymmetric partners are within 1 GeV in mass.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Time-dependent SOLPS-ITER simulations of the tokamak plasma boundary for model predictive control using SINDy *

Abstract Time-dependent SOLPS-ITER simulations have been used to identify reduced models with the sparse identification of nonlinear dynamics (SINDy) method and develop model-predictive control of the boundary plasma state using main ion gas puff actuation. A series of gas actuation sequences are input into SOLPS-ITER to produce a dynamic response in upstream and divertor plasma quantities. The SINDy method is applied to identify reduced linear and nonlinear models for the electron density at the outboard midplane n e , s e p O M P and the electron temperature at the outer divertor T e , s e p d i v . Note that T e , s e p d i v is not necessarily the peak value of T e along the divertor. The identified reduced models are interpretable by construction (i.e. not black box), and have the form of coupled ordinary differential equations. Despite significant noise in T e , s e p d i v , the reduced models can be used to predict the response over a range of actuation levels to a maximum deviation of 0.5% in n e , s e p O M P and 5%–10% in T e , s e p d i v for the cases considered. Model retraining using time history data triggered by a preset error threshold is also demonstrated. A model predictive control strategy for nonlinear models is developed and used to perform feedback control of a SOLPS-ITER simulation to produce a setpoint trajectory in n e , s e p O M P using the integrated plasma simulator framework. The developed techniques are general and can be applied to time-dependent data from other boundary simulations or experimental data. Ongoing work is extending the approach to model identification and control for divertor detachment, which will present transient nonlinear behavior from impurity seeding, including realistic latency and synthetic diagnostic signals derived from the full SOLPS-ITER output.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Benchmark eA generator for leptoproduction in high-energy lepton-nucleus collisions

The upcoming electron-ion collider (EIC) will address several outstanding puzzles in modern nuclear physics. Topics such as the partonic structure of nucleons and nuclei, the origin of their mass and spin, among others, can be understood via the study of high-energy electron-proton (ep) and electron-nucleus (eA) collisions. Achieving the scientific goals of the EIC will require a novel electron-hadron collider and detectors capable to perform high-precision measurements but also dedicated tools to model and interpret the data. To aid in the latter, we present a general-purpose e A Monte Carlo generator—BeAGLE. In this paper, we provide a general description of the models integrated into BeAGLE, applications of BeAGLE in eA physics, implications for detector requirements at the EIC, and the tuning of the parameters in BeAGLE based on available experimental data. Specifically, we focus on a selection of model and data comparisons in particle production in both ep and eA collisions, where baseline particle distributions provide essential information to characterize the event. In addition, we investigate the collision geometry determination in eA collisions, which could be used as an experimental tool for varying the nuclear density.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Search for top squarks in events with a Higgs or Z boson using 139 fb -1 of pp collision data at $\sqrt{s}=13$ TeV with the ATLAS detector

This paper presents a search for direct top squark pair production in events with missing transverse momentum plus either a pair of jets consistent with Standard Model Higgs boson decay into b-quarks or a same-flavour opposite-sign dilepton pair with an invariant mass consistent with a Z boson. The analysis is performed using the proton–proton collision data at √s=13 TeV collected with the ATLAS detector during the LHC Run-2, corresponding to an integrated luminosity of 139 fb -1 . No excess is observed in the data above the Standard Model predictions. The results are interpreted in simplified models featuring direct production of pairs of either the lighter top squark ($\tilde{t}_1$) or the heavier top squark ($\tilde{t}_2$), excluding at 95% confidence level $\tilde{t}_1$ and $\tilde{t}_2$ masses up to about 1220 and 875 GeV, respectively.

42 ENGINEERING↗

Analysis of head-down tilt as an analog of weightlessness using a methematical simulation model

Antiorthostasis or head down tilt of a moderate degree was used as a ground based analog of weightless space flight to study headward fluid shifts, decreased plasma volume, orthostatic intolerance and muscular skeletal degradation. A mathematical model was used to help interpret these observations. The model proved most valuable for these studies was originally developed as a description of the major circulatory, fluid and electrolyte control systems. Two different experimental studies are employed to validate the model. The first is a 24 hour head down tilt study and the second is a 7 day head down bed rest study. The major issues addressed include the reduction in plasma volume, the dynamic changes of venous pressure and cardiac output, the extent of central hypervolemia during long term zero g exposure, the existence of an early diuresis, the mechanisms which alter the renal regulating hormones during the short term and long term periods, the significance of potassium loss on other zero g responses, and the role of transcapillary filtration in adjusting fluid shifts. The use of mathematical models as an interpretive and analysis technique for experimental research for space life science is illustrated.

Leonard, J. I.↗