Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “model selection”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

Future Projections of Lifecycle Cost and Greenhouse Gas Emissions of Light-Duty Vehicles

Vehicles with electrified powertrains carry the promise of significant reductions in greenhouse gas (GHG) emissions from a lifecycle analysis (LCA) standpoint compared to conventional internal combustion engine (CICE) vehicles. However, trade-offs exist between different types of electrified powertrains in terms of cost, consumer acceptance, and GHG reduction efficacy for different operating conditions. The open-source tool CarGHG was developed with an aim to enable the exploration of a plethora of parametric study scenarios, including the cost of electrification technologies, different driving patterns and charging habits, and the cost and carbon intensity of electricity and fuel blends. This paper introduces the framework of CarGHG, then showcases total cost of ownership (TCO) and LCA GHG results for select models of light-duty vehicles. Another capability of CarGHG, which is the ability to estimate the performance of “virtual” vehicle models (perceived vehicle design specifications not yet on the market), is utilized to explore future scenarios of electrification and low-carbon fuel blends for Small Sports Utility Vehicles (SUVs), a popular light-duty vehicle segment in North America. With opportunities, but also uncertainties, in future scenarios, it is likely wise to continue pursuing multiple ways towards the reduction of LCA GHG.

Hamza, Karim↗

PaleoSTeHM v1.0: a modern, scalable spatiotemporal hierarchical modeling framework for paleo-environmental data

Abstract. Geological records of past environmental change provide crucial insights into long-term climate variability, trends, non-stationarity, and nonlinear feedback mechanisms. However, reconstructing spatiotemporal fields from these records is statistically challenging due to their sparse, indirect, and noisy nature. Here, we present PaleoSTeHM, a scalable and modern framework for spatiotemporal hierarchical modeling of paleo-environmental data. This framework enables the implementation of flexible statistical models that rigorously quantify spatial and temporal variability from geological data while clearly distinguishing measurement and inferential uncertainty from process variability. We illustrate its application by reconstructing temporal and spatiotemporal paleo-sea-level changes across multiple locations. Using various modeling and analysis choices, PaleoSTeHM demonstrates the impact of different methods on inference results and computational efficiency. Our results highlight the critical role of model selection in addressing specific paleo-environmental questions, showcasing the PaleoSTeHM framework's potential to enhance the robustness and transparency of paleo-environmental reconstructions.

58 GEOSCIENCES↗

Beyond pinball loss: Quantile methods for calibrated uncertainty quantification

Amongthemanywaysofquantifying uncertainty in a regression setting, specifying the full quantile function is attractive, as quantiles are amenable to interpretation and evaluation. A model that predicts the true conditional quantiles for each input, at all quantile levels, presents a correct and efficient representation of the underlying uncertainty. To achieve this, many current quantile-based methods focus on optimizing the pinball loss. However, this loss restricts the scope of applicable regression models, limits the ability to target many desirable properties (e.g. calibration, sharpness, centered intervals), and may produce poor conditional quantiles. In this work, we develop new quantile methods that address these shortcomings. In particular, we propose methods that can apply to any class of regression model, select an explicit balance between calibration and sharpness, optimize for calibration of centered intervals, and produce more accurate conditional quantiles. We provide a thorough experimental evaluation of our methods, which includes a high dimensional uncertainty quantification task in nuclear fusion.

97 MATHEMATICS AND COMPUTING↗

Downscaled Daily 1 km Climate Data (NEX-GDDP-CMIP6) for Southeast Texas. Full ensemble of downscaled CMIP6 climate projections at 1 km daily resolution.

For the SETx-UIFL, the daily NASA Earth Exchange Global Daily Downscaled Projections (NEX-GDDP-CMIP6) dataset climate projections were downscaled from approximately 27 km to 1 km. The SETx dataset provides very high-resolution climate data for the historical period (1950–2014) and future scenarios derived from CMIP6 global models under the four Tier 1 Shared Socioeconomic Pathways (SSPs 1.26, 2.45, 3.70, and 5.85), developed for the IPCC Sixth Assessment Report. A subset of ten NEX-GDDP-CMIP6 models was selected to represent a balance of model families, climate sensitivities, and availability across scenarios, ensuring a diverse and reliable ensemble for regional analysis. Selected models: BCC-CSM2-MR, CESM2, CMCC-ESM2, CNRM-ESM2-1, EC-Earth3, FGOALS-g3, GFDL-CM4, MPI-ESM1-2-HR, MRI-ESM2-0, NorESM2-MM. Daily variables downscaled include tasmax, tasmin, tas, pr, hurs, huss, rsds, rlds, and sfcWind.

Persad, Geeta↗

Signal selection and model-independent extraction of the neutrino neutral-current single 𝜋 + cross section with the T2K experiment

This article presents a study of single 𝜋 + production in neutrino neutral-current interactions (NC⁢1⁢𝜋 + ) using the FGD1 hydrocarbon target of the ND280 detector of the T2K experiment. We report the largest sample of such events selected by any experiment, providing the first new data for this channel in over four decades and the first using a sub-GeV neutrino flux. The signal selection strategy and its performance are detailed together with validations of a robust cross section extraction methodology. The measured flux-averaged integrated cross-section is 𝜎 = (6.07 ± 1.22) × 10 −41 cm 2 /nucleon, 1.3⁢𝜎 above the NEUT v5.4.0 expectation.

Neutrino detection↗

FluxRETAP: a REaction TArget Prioritization genome-scale modeling technique for selecting genetic targets

MOTIVATION: Metabolic engineering is rapidly evolving as a result of new advances in synthetic biology tools and automation platforms that enable high throughput strain construction, as well as the development of machine learning tools (ML) for biology. However, selecting genetic engineering targets that effectively guide the metabolic engineering process is still challenging. ML can provide predictive power for synthetic biology, but current technical limitations prevent the independent use of ML approaches without previous biological knowledge. RESULTS: Here, we present FluxRETAP, a simple and computationally inexpensive method that leverages the prior mechanistic knowledge embedded in genome-scale models for suggesting targets for genetic overexpression, downregulation or deletion, with the final goal of increasing the production of a desired metabolite. This method can provide a list of desirable engineering targets that can be combined with current ML pipelines. FluxRETAP captured 100% of reaction targets experimentally verified to improve Escherichia coli isoprenol production, 50% of targets that experimentally improved taxadiene production in E. coli and ∼60% of genetic targets from a verified minimal constrained cut-set in Pseudomonas putida, while providing additional high priority targets that could be tested. Overall, FluxRETAP is an efficient algorithm for identifying a prioritized list of testable genetic and reaction targets. AVAILABILITY AND IMPLEMENTATION: FluxRETAP is implemented in python and released under the creative commons license. The implementation and code are freely available at: https://github.com/JBEI/FluxRETAP.

Czajka, Jeffrey J↗

Informing ARC divertor design and plasma facing material selection through integrated modeling

This INFUSE 2023 project between UCLA and Commonwealth Fusion Systems used computer modeling to test whether tungsten materials can survive in CFS's ARC fusion reactor divertor. The team simulated plasma conditions and material responses, finding that tungsten-rhenium alloys resist grain growth better than pure tungsten, and that hydrogen buildup depends more on particle flux than temperature. The work helps CFS design durable plasma-facing components for their fusion power plant.

36 MATERIALS SCIENCE↗

Selection of Global Climate Model Data for Downscaling With Generative Machine Learning and Use in the Power Planning for Alignment of Climate and Energy Systems Project

The range of results from climate models and scenarios is important to the understanding of uncertainty in power planning analysis. A U.S. Department of Energy-funded analytic project called Power Planning for Alignment of Climate and Energy Systems is developing data and analytic methods to reflect the effects of climate change on key variables for power system planning, as part of the Grid Modernization Lab Consortium. This project will select and prepare global climate model results for use in power system planning models. A related report (Evaluation of Global Climate Models for Use in Energy Analysis) assesses the performance of various global climate models from the Coupled Model Intercomparison Project Phase 6 data archive for their historical skill with respect to energy system performance and for their future projections under multiple climate change scenarios. Building from that report, we describe the selection of a climate scenario (Shared Socioeconomic Pathway [SSP] 2-4.5) and five climate models: TaiESM1, EC-Earth3-CC, GFDL-CM4, EC-Earth3-Veg, and MPI-ESM1-2-HR. We describe the model selection criteria, which were based on the quality of the match between model results under historical conditions and on the representation of the range of future values for several variables. These results will be downscaled via an open-source generative machine learning method called Super-Resolution for Renewable Energy Resource Data with Climate Change Impacts.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Review of data-driven models for quantifying load shed by non-residential buildings in the United States

Shifting and shedding power demand in buildings can be cost-effective techniques for grids to function reliably and for end users to earn compensation. Grid operators reimburse customers in proportion to the quantity of load shed. Simple data-driven methods are used to quantify this shed, which is the difference between a measured load during the event and modeled "baseline" that would have occurred in absence of the event. These methods have evolved over the years and in many cases have been integrated with building physics, to make them a hybrid between physics based and empirical models. However, there is no comprehensive analysis that provides guidance to building operators, grid operators and researchers in selecting appropriate models based on their specific needs and available data. Here, this work aims to fill this gap by critically assessing the performance of baseline models put forward from the year 2000 through 2023. The literature reviewed includes reports generated by grid operators, reports from national laboratories and academic journal articles. The work outlines modeling features like the inputs, training period, estimation method, adjustments to fine tune the predictions and metrics to evaluate the performance. A comprehensive list of 50 models has been provided. For each model, the study explores the applicability of the model to weather sensitive buildings, variability in the building profile, timing of the event, and whether the building reduces energy consumption before an event. The work identifies the situations in which a particular model works and draws lessons based on evidence of performance. Finally, recommendations to aid in model selection are given.

97 MATHEMATICS AND COMPUTING↗

Modeling and Calibration of Supplier Selection Problem in Freight Agent-Based Simulations

Freight transportation modeling often struggles with data limitations, especially in accurately representing complex supplier selection processes and their impact on network flows. This research addresses this critical gap by developing a large-scale, calibrated agent-based model for supplier selection, complemented by a probabilistic heuristic for international shipments. Our approach integrates trade relationships between industry sectors, transportation costs, and a supplier-rating model adapted from existing literature. The model’s core objective is to minimize the discrepancy between modeled and observed commodity flows while ensuring a close match to regional shipping distance distributions. Implemented and tested across four major U.S. metropolitan areas—Atlanta, Chicago, Dallas–Fort Worth, and Los Angeles—the model demonstrates high fidelity in replicating observed freight patterns. Key findings reveal consistent alignment with national shipping distance trends and highlight significant spatial variations in commodity trade assignments and demand across the study regions. This behaviorally informed and transport-sensitive framework is designed to approximate real-world decision making, providing a robust tool for policymakers and planners to evaluate targeted interventions, assess infrastructure investments, and enhance supply chain resilience in the face of disruptions.

Ismael, Abdelrahman (ORCID:0000000303712110)↗

Correcting for Selection Biases in the Determination of the Hubble Constant from Time-Delay Cosmography

The time delay between multiple images of strongly lensed quasars has been used to infer the Hubble constant. The primary systematic uncertainty for time-delay cosmography is the mass-sheet transform (MST), which preserves the lensing observables while altering the inferred ⁠H 0 . The TDCOSMO collaboration used velocity dispersion measurements of lensed quasars and lensed galaxies to infer that mass sheets are present, which decrease the inferred H 0 by 8 per cent. Here, we test the assumption that the density profiles of galaxy–galaxy and galaxy–quasar lenses are the same. We use a composite star-plus-dark-matter mass profile for the parent deflector population and model the selection function for galaxy–galaxy and galaxy–quasar lenses. We find that a power-law density profile with an MST is a good approximation to a two-component mass profile around the Einstein radius, but we find that galaxy–galaxy lenses have systematically higher mass-sheet components than galaxy–quasar lenses. For individual systems, λ int correlates with the ratio of the half-light radius and Einstein radius of the lens. By propagating these results through the TDCOSMO hierarchical inference code, we find that H 0 is lowered by a further 3 per cent. Using a more recent measurement of velocity dispersions and our fiducial model for selection biases, we infer H 0 = 66 ± 4 (stat) ± 1 (model sys) ± 2 (measurement sys) km s -1 Mpc -1 for the TDCOSMO plus SLACS data set. The first residual systematic error is due to plausible alternative choices in modelling the selection function, and the second is an estimate of the remaining systematic error in the measurement of velocity dispersions for SLACS lenses. Accurate time-delay cosmography requires precise velocity dispersion measurements and accurate calibration of selection biases.

79 ASTRONOMY AND ASTROPHYSICS↗

Selection function of clusters in Dark Energy Survey year 3 data from cross-matching with South Pole Telescope detections

Context. Galaxy clusters selected based on overdensities of galaxies in photometric surveys provide the largest cluster samples. However, modeling the selection function of such samples is complicated by noncluster members projected along the line of sight (projection effects) and the potential detection of unvirialized objects (contamination). Aims. We empirically constrained the magnitude of these effects by cross-matching galaxy clusters selected in the Dark Energy Survey data with the redMaPPer algorithm with significant detections in three South Pole Telescope surveys (SZ, pol-ECS, pol-500d). Methods. For matched clusters, we augmented the redMaPPer catalog with the SPT detection significance. For unmatched objects we used the SPT detection threshold as an upper limit on the SZe signature. Using a Bayesian population model applied to the collected multiwavelength data, we explored various physically motivated models to describe the relationship between observed richness and halo mass. Results. Our analysis reveals a clear preference for models with an additional skewed scatter component associated with projection effects over a purely log-normal scatter model. We rule out significant contamination by unvirialized objects at the high-richness end of the sample. While dedicated simulations offer a well-fitting calibration of projection effects, our findings suggest the presence of redshift-dependent trends that these simulations may not have captured. Our findings highlight that modeling the selection function of optically detected clusters remains a complicated challenge that requires a combination of simulation and data-driven approaches.

79 ASTRONOMY AND ASTROPHYSICS↗

Machine learning approaches for influenza A virus risk assessment identifies predictive correlates using ferret model in vivo data

In vivo assessments of influenza A virus (IAV) pathogenicity and transmissibility in ferrets represent a crucial component of many pandemic risk assessment rubrics, but few systematic efforts to identify which data from in vivo experimentation are most useful for predicting pathogenesis and transmission outcomes have been conducted. To this aim, we aggregated viral and molecular data from 125 contemporary IAV (H1, H2, H3, H5, H7, and H9 subtypes) evaluated in ferrets under a consistent protocol. Three overarching predictive classification outcomes (lethality, morbidity, transmissibility) were constructed using machine learning (ML) techniques, employing datasets emphasizing virological and clinical parameters from inoculated ferrets, limited to viral sequence-based information, or combining both data types. Among 11 different ML algorithms tested and assessed, gradient boosting machines and random forest algorithms yielded the highest performance, with models for lethality and transmission consistently better performing than models predicting morbidity. Comparisons of feature selection among models was performed, and highest performing models were validated with results from external risk assessment studies. Our findings show that ML algorithms can be used to summarize complex in vivo experimental work into succinct summaries that inform and enhance risk assessment criteria for pandemic preparedness that take in vivo data into account.

59 BASIC BIOLOGICAL SCIENCES↗

Optimization-Based Dynamic Voltage Support of Microgrids Using Energy Storage Systems

A microgrid network is characterized by a high R/X ratio, making the voltage more sensitive to active power changes compared to bulk power systems, where the voltage is regulated primarily by reactive power. Due to its sensitivity, voltage control approaches for microgrids should also consider the active power input coupling, making it very different from conventional power systems. Additionally, as the energy costs associated with active and reactive powers are different and the operational conditions of microgrids connected to active distribution systems vary over time, the ideal controller to provide voltage support must be flexible enough to handle these technical and operational constraints. This paper proposes a model predictive control approach to provide dynamic voltage support using energy storage systems. This approach uses a simplified predictive model of the system to solve the model predictive control problem. By proper selection of model predictive control weighting parameters, the quality of service provided can be adjusted to achieve the desired performance. A simulation study in MATLAB/Simulink validates the proposed approach for the Cordova, Alaska microgrid. Results show that the performance of the voltage support can be adjusted depending on the choice of weight and constraints of the controller.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Model Calibration with Markov Chain Monte Carlo Tutorial

The purpose of this tutorial is to demonstrate how to use Markov chain Monte Carlo (MCMC) to calibrate a model. By calibration, we mean the selection of model parameters (and, when relevant, structures). A common goal in model development and diagnostics is calibration, or the identification of model structures and parameters which are consistent with data. While models can be calibrated through hand-tuning parameters or minimizing simple error metrics such as root-mean-square-error (RMSE), these approaches can underrepresent the probabilistic nature of the data-generating process, as well as the potential for multiple model configurations to be consistent with the data. Probabilistic uncertainty quantification, which is the topic of this notebook, can address these concerns. This tutorial is presented as an appendix to the e-book: Addressing Uncertainty in MultiSector Dynamics Research.

Markov chain Monte Carlo↗

Material Selection and Heat Transfer Model for PTUHS Device

The Path to Ultimate Heat Sink Device (PTUHS) is a passive safety system that utilizes a radiative heat transfer valve to control the rate at which heat is removed from a nuclear reactor pressure vessel (RPV) and dispersed into surrounding soil. The material selection for the PTUHS device is investigated, where the thermophysical properties are either maximized or minimized bases on what section of the PTUHS that material is being used for. Final recommendations of material choice are then presented. The heat transfer model of the PTUHS device is solved in ABAQUS, where a normal operating conditions and SCRAM conditions are solved. These temperature maps and heat flux of the system show the system's ability, where minimal heat is lost during normal operating conditions and system failure in a SCRAM scenario due to temperatures that were reached.

42 - ENGINEERING↗

Model-independent predictions for decays of hidden-heavy hadrons into pairs of heavy hadrons

Hidden-heavy hadrons can decay into pairs of heavy hadrons through transitions from confining Born-Oppenheimer potentials to hadron-pair potentials with the same Born-Oppenheimer quantum numbers. The transitions are also constrained by conservation of angular momentum and parity. From these constraints, we derive model-independent selection rules for decays of hidden-heavy hadrons into pairs of heavy hadrons. The coupling potentials are expressed as sums of products of Born-Oppenheimer transition amplitudes and angular-momentum coefficients. If there is a single dominant Born-Oppenheimer transition amplitude, it factors out of the coupling potentials between hidden-heavy hadrons in the same Born-Oppenheimer multiplet and pairs of heavy hadrons in specific heavy-quark-spin-symmetry doublets. If furthermore the kinetic energies of the heavy hadrons are much larger than their spin splittings, we obtain analytic expressions for the relative partial decay rates in terms of Wigner 6 j and 9 j symbols. We consider in detail the decays of quarkonia and quarkonium hybrids into the lightest heavy-meson pairs. For quarkonia, our model-independent selection rules and relative partial decay rates agree with previous results from quark-pair-creation models in simple cases and give stronger results in other cases. For quarkonium hybrids, we find disagreement even in simple cases. Published by the American Physical Society 2024

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Part-scale microstructure prediction for laser powder bed fusion Ti-6Al-4V using a hybrid mechanistic and machine learning model

Laser powder bed fusion (LPBF) Ti-6Al-4V is widely studied for use in structural applications in aerospace and medical industries, but mechanical anisotropy and microstructural inhomogeneity prohibits its wider adoption. Although successful microstructure prediction models have been developed, a remaining challenge is their limited integration across length/time scales and validation by experimental studies. Here, this work proposes a physics-augmented machine learning surrogate model to unite predictions of LPBF temperature, β phase morphology and texture, and α/α’ formation into a single framework that is calibrated and validated with experiments. First, a phase field (PF) model of the martensitic β→α’ transformation is developed and calibrated using data from in-situ synchrotron cyclic heating/cooling studies quantifying the variation of α phase fraction with time. In parallel, an established finite difference-Monte Carlo (FDMC) model predicts the part-scale temperature profile and β grain formation during solidification. A dataset is developed using LPBF cyclic temperature descriptors from the FDMC model as inputs and corresponding α/α’ phase fraction and width from the PF model as outputs. Five machine learning (ML) regression models are tested and optimized, having mean absolute error in testing ≤ 4 %, and the k-nearest neighbors (KNN) model is selected as the best performing. The KNN model is called at the nodal level during post-processing of the FDMC model to replace and downscale the response of the PF model. The combined agility and accuracy of the hybrid FDMC-ML model enables part-scale microstructure predictions that can be further used for property predictions to accelerate AM process optimization.

36 MATERIALS SCIENCE↗