Engineering PapersSearch

SEARCH · Engineering Papers

Results for “ensemble”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Producing High-fidelity Synthetic Population Ensembles at Scale

Used within social simulations, synthetic population ensembles enable uncertainty quantification (UQ) methods for obtaining more robust model inference and prediction. A synthetic population ensemble is a series of plausible virtual reconstructions of an area’s population at the granularity of people and residences, generated stochastically to preserve privacy of the source population survey’s respondents. In this paper, we demonstrate the production of large synthetic population ensembles for the US via Oak Ridge National Laboratory’s UrbanPop framework to support modeling of high spatial resolution energy affordability metrics from nationwide social surveys in collaboration with the fusionACS project. Our initial task involves creating ensembles for 17 US metropolitan areas, each consisting of 41 population instances (a base realization and 40 replicates). To accomplish this task at scale, we configured an integrated system comprised of a research cloud, virtual containerization, GPU-enhanced functionality, and a dual API/CLI to interact with UrbanPop’s maturing Likeness Python ecosystem. We observe a reduction in theoretical execution time while maintaining high-fidelity approximations of residential totals by metropolitan area and the demographic characteristics of neighborhoods. We discuss expansion of our approach to produce synthetic population ensembles for the entire US, particularly plans to establish automated workflows for job orchestration to increase computational efficiency, as well as provide outlook for broadening applications of the ensembles.

Gaboardi, James [ORNL] (ORCID:0000000247766826)

New Software for Ensemble Creation in the Spitzer-Space-Telescope Operations Database

Some of the computer pipelines used to process digital astronomical images from NASA's Spitzer Space Telescope require multiple input images, in order to generate high-level science and calibration products. The images are grouped into ensembles according to well documented ensemble-creation rules by making explicit associations in the operations Informix database at the Spitzer Science Center (SSC). The advantage of this approach is that a simple database query can retrieve the required ensemble of pipeline input images. New and improved software for ensemble creation has been developed. The new software is much faster than the existing software because it uses pre-compiled database stored-procedures written in Informix SPL (SQL programming language). The new software is also more flexible because the ensemble creation rules are now stored in and read from newly defined database tables. This table-driven approach was implemented so that ensemble rules can be inserted, updated, or deleted without modifying software.

Spitzer Science Center (SSC)

An Ensemble-Based Smoother with Retrospectively Updated Weights for Highly Nonlinear Systems

Monte Carlo computational methods have been introduced into data assimilation for nonlinear systems in order to alleviate the computational burden of updating and propagating the full probability distribution. By propagating an ensemble of representative states, algorithms like the ensemble Kalman filter (EnKF) and the resampled particle filter (RPF) rely on the existing modeling infrastructure to approximate the distribution based on the evolution of this ensemble. This work presents an ensemble-based smoother that is applicable to the Monte Carlo filtering schemes like EnKF and RPF. At the minor cost of retrospectively updating a set of weights for ensemble members, this smoother has demonstrated superior capabilities in state tracking for two highly nonlinear problems: the double-well potential and trivariate Lorenz systems. The algorithm does not require retrospective adaptation of the ensemble members themselves, and it is thus suited to a streaming operational mode. The accuracy of the proposed backward-update scheme in estimating non-Gaussian distributions is evaluated by comparison to the more accurate estimates provided by a Markov chain Monte Carlo algorithm.

Monte Carlo

Improving Climate Projections Using "Intelligent" Ensembles

Recent changes in the climate system have led to growing concern, especially in communities which are highly vulnerable to resource shortages and weather extremes. There is an urgent need for better climate information to develop solutions and strategies for adapting to a changing climate. Climate models provide excellent tools for studying the current state of climate and making future projections. However, these models are subject to biases created by structural uncertainties. Performance metrics-or the systematic determination of model biases-succinctly quantify aspects of climate model behavior. Efforts to standardize climate model experiments and collect simulation data-such as the Coupled Model Intercomparison Project (CMIP)-provide the means to directly compare and assess model performance. Performance metrics have been used to show that some models reproduce present-day climate better than others. Simulation data from multiple models are often used to add value to projections by creating a consensus projection from the model ensemble, in which each model is given an equal weight. It has been shown that the ensemble mean generally outperforms any single model. It is possible to use unequal weights to produce ensemble means, in which models are weighted based on performance (called "intelligent" ensembles). Can performance metrics be used to improve climate projections? Previous work introduced a framework for comparing the utility of model performance metrics, showing that the best metrics are related to the variance of top-of-atmosphere outgoing longwave radiation. These metrics improve present-day climate simulations of Earth's energy budget using the "intelligent" ensemble method. The current project identifies several approaches for testing whether performance metrics can be applied to future simulations to create "intelligent" ensemble-mean climate projections. It is shown that certain performance metrics test key climate processes in the models, and that these metrics can be used to evaluate model quality in both current and future climate states. This information will be used to produce new consensus projections and provide communities with improved climate projections for urgent decision-making.

Baker, Noel C.

Improving Climate Projections Using "Intelligent" Ensembles

Recent changes in the climate system have led to growing concern, especially in communities which are highly vulnerable to resource shortages and weather extremes. There is an urgent need for better climate information to develop solutions and strategies for adapting to a changing climate. Climate models provide excellent tools for studying the current state of climate and making future projections. However, these models are subject to biases created by structural uncertainties. Performance metrics-or the systematic determination of model biases-succinctly quantify aspects of climate model behavior. Efforts to standardize climate model experiments and collect simulation data-such as the Coupled Model Intercomparison Project (CMIP)-provide the means to directly compare and assess model performance. Performance metrics have been used to show that some models reproduce present-day climate better than others. Simulation data from multiple models are often used to add value to projections by creating a consensus projection from the model ensemble, in which each model is given an equal weight. It has been shown that the ensemble mean generally outperforms any single model. It is possible to use unequal weights to produce ensemble means, in which models are weighted based on performance (called "intelligent" ensembles). Can performance metrics be used to improve climate projections? Previous work introduced a framework for comparing the utility of model performance metrics, showing that the best metrics are related to the variance of top-of-atmosphere outgoing longwave radiation. These metrics improve present-day climate simulations of Earth's energy budget using the "intelligent" ensemble method. The current project identifies several approaches for testing whether performance metrics can be applied to future simulations to create "intelligent" ensemble-mean climate projections. It is shown that certain performance metrics test key climate processes in the models, and that these metrics can be used to evaluate model quality in both current and future climate states. This information will be used to produce new consensus projections and provide communities with improved climate projections for urgent decision-making.

Baker, Noel C.

ESD Reviews: Model Dependence in Multi-Model Climate Ensembles: Weighting, Sub-Selection and Out-Of-Sample Testing

The rationale for using multi-model ensembles in climate change projections and impacts research is often based on the expectation that different models constitute independent estimates; therefore, a range of models allows a better characterisation of the uncertainties in the representation of the climate system than a single model. However, it is known that research groups share literature, ideas for representations of processes, parameterisations, evaluation data sets and even sections of model code. Thus, nominally different models might have similar biases because of similarities in the way they represent a subset of processes, or even be near-duplicates of others, weakening the assumption that they constitute independent estimates. If there are near-replicates of some models, then treating all models equally is likely to bias the inferences made using these ensembles. The challenge is to establish the degree to which this might be true for any given application. While this issue is recognised by many in the community, quantifying and accounting for model dependence in anything other than an ad-hoc way is challenging. Here we present a synthesis of the range of disparate attempts to define, quantify and address model dependence in multi-model climate ensembles in a common conceptual framework, and provide guidance on how users can test the efficacy of approaches that move beyond the equally weighted ensemble. In the upcoming Coupled Model Intercomparison Project phase 6 (CMIP6), several new models that are closely related to existing models are anticipated, as well as large ensembles from some models. We argue that quantitatively accounting for dependence in addition to model performance, and thoroughly testing the effectiveness of the approach used will be key to a sound interpretation of the CMIP ensembles in future scientific studies.

Abramowitz, Gab

Uncertainty in Soil Moisture Retrievals: an Ensemble Approach Using SMOS L-Band Microwave Data

The uncertainty of soil moisture (SM) retrievals from satellite brightness temperature (TB) observations depends primarily on the choice of radiative transfer model (RTM) parameters, prior SM information and TB inputs. This paper studies the sensitivity of several (quasi-)operational and experimental SM retrieval products from the Soil Moisture Ocean Salinity (SMOS) mission to these choices at 11 reference sites, located in 7 watersheds across the United States (US). Different literature-based RTM parameter sets cause large biases between retrievals. Whereas typical RTM parameter sets are calibrated for SM retrievals, it is shown that a parameter set carefully optimized for TB forward modeling can also be used for retrieving SM. It is also shown that the inclusion of dynamic prior SM estimates in a Bayesian retrieval scheme can strongly improve SM retrievals, regardless of the choice of RTM parameters, and that the use of multi-angular and multi-polarization TB does not necessarily lead to superior retrievals compared to retrievals based on TB data at a single incidence angle and polarization. The second part of this paper evaluates ensemble uncertainty metrics for SM retrievals obtained by propagating a wide range of RTM parameters through the RTM. As expected for bounded variables, the spread in the ensemble SM retrievals is smallest for wet and dry SM values and highest for intermediate SM values. After removal of the strong long-term mean bias associated with the RTM parameter values for individual ensemble members, the remaining anomaly ensemble SM spread of 0.037 cu m/cu m approximates the actual time series unbiased root-mean-square-difference of 0.042 cu m/cu m between ensemble mean retrievals and in situ reference data across the reference sites. However, the temporal variability in the anomaly ensemble spread reveals higher-order biases in the retrieval error, which should be accounted for when characterizing retrieval error.

Jan Quets

Parameterization-Induced Uncertainties and Impacts of Crop Management Harmonization in a Global Gridded Crop Model Ensemble

Global gridded crop models (GGCMs) combine agronomic or plant growth models with gridded spatial input data to estimate spatially explicit crop yields and agricultural externalities at the global scale. Differences in GGCM outputs arise from the use of different biophysical models, setups, and input data. GGCM ensembles are frequently employed to bracket uncertainties in impact studies without investigating the causes of divergence in outputs. This study explores differences in maize yield estimates from five GGCMs based on the public domain field-scale model Environmental Policy Integrated Climate (EPIC) that participate in the AgMIP Global Gridded Crop Model Intercomparison initiative. Albeit using the same crop model, the GGCMs differ in model version, input data, management assumptions, parameterization, and selection of subroutines affecting crop yield estimates via cultivar distributions, soil attributes, and hydrology among others. The analyses reveal inter-annual yield variability and absolute yield levels in the EPIC-based GGCMs to be highly sensitive to soil parameterization and crop management. All GGCMs show an intermediate performance in reproducing reported yields with a higher skill if a static soil profile is assumed or sufficient plant nutrients are supplied. An in-depth comparison of setup domains for two EPIC-based GGCMs shows that GGCM performance and plant stress responses depend substantially on soil parameters and soil process parameterization, i.e. hydrology and nutrient turnover, indicating that these often neglected domains deserve more scrutiny. For agricultural impact assessments, employing a GGCM ensemble with its widely varying assumptions in setups appears the best solution for coping with uncertainties from lack of comprehensive global data on crop management, cultivar distributions and coefficients for agro-environmental processes. However, the underlying assumptions require systematic specifications to cover representative agricultural systems and environmental conditions. Furthermore, the interlinkage of parameter sensitivity from various domains such as soil parameters, nutrient turnover coefficients, and cultivar specifications highlights that global sensitivity analyses and calibration need to be performed in an integrated manner to avoid bias resulting from disregarded core model domains. Finally, relating evaluations of the EPIC-based GGCMs to a wider ensemble based on individual core models shows that structural differences outweigh in general differences in configurations of GGCMs based on the same model, and that the ensemble mean gains higher skill from the inclusion of structurally different GGCMs. Although the members of the wider ensemble herein do not consider crop-soil-management interactions, their sensitivity to nutrient supply indicates that findings for the EPIC-based sub-ensemble will likely become relevant for other GGCMs with the progressing inclusion of such processes.

Folberth, Christian

A NASA GISTEMP Observational Uncertainty Ensemble: Regional and Monthly Uncertainty

The historical global temperature record is an essential data product for quantifying the variability and change of the Earth system. In recent years, better characterization of observational uncertainty in global and hemispheric trends has become available, but the methodologies are not necessarily applicable to analyses at smaller regional areas, or monthly means, where station sparsity and other systematic issues contribute to greater uncertainty. This work details a gridded uncertainty ensemble of historical temperature anomalies from the Goddard Institute for Space Studies (GISS) Surface Temperature product (GISTEMP) product. This ensemble characterizes the complex spatial and temporal correlation structure of uncertainty in gridded historical temperature, enabling proper uncertainty propagation for climate and social science at regional and monthly scales. This work details the methodology for generating the uncertainty ensemble, key statistics of the uncertainty evolution over space and time, and provides best practices for using the uncertainty ensemble in future studies. Summary statistics from the uncertainty ensemble are in good agreement with production GISTEMP. Two applications of the uncertainty ensemble are also presented. First, the warmest year on record is shown to most likely be 2016 with a 53.2% chance and 2020 as the second most likely with a 44.4% chance. Second, it is shown that the arctic is warming 2.5 - 5 times faster than the globe, significantly faster than the regularly quoted twice as fast.

GISTEMP

The optimization of model ensemble composition and size can enhance the robustness of crop yield projections

Linked climate and crop simulation models are widely used to assess the impact of climate change on agriculture. However, it is unclear how ensemble configurations (model composition and size) influence crop yield projections and uncertainty. Here, we investigate the influences of ensemble configurations on crop yield projections and modeling uncertainty from Global Gridded Crop Models and Global Climate Models under future climate change. We performed a cluster analysis to identify distinct groups of ensemble members based on their projected outcomes, revealing unique patterns in crop yield projections and corresponding uncertainty levels, particularly for wheat and soybean. Furthermore, our findings suggest that approximately six Global Gridded Crop Models and 10 Global Climate Models are sufficient to capture modeling uncertainty, while a cluster-based selection of 3-4 Global Gridded Crop Models effectively represents the full ensemble. The contribution of individual Global Gridded Crop Models to overall uncertainty varies depending on region and crop type, emphasizing the importance of considering the impact of specific models when selecting models for local-scale applications. Our results emphasize the importance of model composition and ensemble size in identifying the primary sources of uncertainty in crop yield projections, offering valuable guidance for optimizing ensemble configurations in climate-crop modeling studies tailored to specific applications.

Agriculture

Scalable Generation of High-fidelity Synthetic Population Ensembles

Used within social simulations, synthetic population ensembles enable uncertainty quantification (UQ) methods for obtaining more robust model inference and prediction. A synthetic population ensemble is a series of plausible virtual reconstructions of an area’s population at the granularity of people and residences, generated stochastically to preserve privacy of the source population survey’s respondents. In this paper, we demonstrate the production of large synthetic population ensembles for the U.S. via Oak Ridge National Laboratory’s UrbanPop framework to support modeling of high spatial resolution energy affordability metrics from nationwide social surveys in collaboration with the fusionACS project. The study involves two scenarios: creating ensembles for (1) 17 U.S. metropolitan areas in 2019 and (2) full U.S. Census Divisions in 2023, with each scenario consisting of 41 population instances (a base realization and 40 replicates). To accomplish this task at scale, we configured an integrated system within a research cloud, comprised of virtual containerizations, GPU-enhanced functionality, and orchestrated deployments of UrbanPop’s maturing Likeness Python ecosystem. Results demonstrate we maintained high-fidelity approximations of residential totals by areas of interest and the demographic characteristics of neighborhoods while reducing manual workflow burdens. Finally, we discuss plans to fine-tune and further develop our automated workflows for truly distributed job orchestration to increase computational efficiency, as well as provide an outlook for broadening applications of the ensembles.

Cluster computing

Minimizing CGYRO HPC Communication Costs in Ensembles with XGYRO by Sharing the Collisional Constant Tensor Structure

First-principles fusion plasma simulations are both compute and memory intensive, and CGYRO is no exception. The use of many HPC nodes to fit the problem in the available memory thus results in significant communication overhead, which is hard to avoid for any single simulation. That said, most fusion studies are composed of ensembles of simulations, so we developed a new tool, named XGYRO, that executes a whole ensemble of CGYRO simulations as a single HPC job. By treating the ensemble as a unit, XGYRO can alter the global buffer distribution logic and apply optimizations that are not feasible on any single simulation, but only on the ensemble as a whole. The main saving comes from the sharing of the collisional constant tensor structure, since its values are typically identical between parameter-sweep simulations. This data structure dominates the memory consumption of CGYRO simulations, so distributing it among the whole ensemble results in drastic memory savings for each simulation, which in turn results in overall lower communication overhead.

CGYRO

Multimodel Ensembles of Wheat Growth: More Models are Better than One

Crop models of crop growth are increasingly used to quantify the impact of global changes due to climate or crop management. Therefore, accuracy of simulation results is a major concern. Studies with ensembles of crop models can give valuable information about model accuracy and uncertainty, but such studies are difficult to organize and have only recently begun. We report on the largest ensemble study to date, of 27 wheat models tested in four contrasting locations for their accuracy in simulating multiple crop growth and yield variables. The relative error averaged over models was 24-38% for the different end-of-season variables including grain yield (GY) and grain protein concentration (GPC). There was little relation between error of a model for GY or GPC and error for in-season variables. Thus, most models did not arrive at accurate simulations of GY and GPC by accurately simulating preceding growth dynamics. Ensemble simulations, taking either the mean (e-mean) or median (e-median) of simulated values, gave better estimates than any individual model when all variables were considered. Compared to individual models, e-median ranked first in simulating measured GY and third in GPC. The error of e-mean and e-median declined with an increasing number of ensemble members, with little decrease beyond 10 models. We conclude that multimodel ensembles can be used to create new estimators with improved accuracy and consistency in simulating growth dynamics. We argue that these results are applicable to other crop species, and hypothesize that they apply more generally to ecological system models.

wheat

Multimodel Ensembles of Wheat Growth: Many Models are Better than One

Crop models of crop growth are increasingly used to quantify the impact of global changes due to climate or crop management. Therefore, accuracy of simulation results is a major concern. Studies with ensembles of crop model scan give valuable information about model accuracy and uncertainty, but such studies are difficult to organize and have only recently begun. We report on the largest ensemble study to date, of 27 wheat models tested in four contrasting locations for their accuracy in simulating multiple crop growth and yield variables. The relative error averaged over models was 2438 for the different end-of-season variables including grain yield (GY) and grain protein concentration (GPC). There was little relation between error of a model for GY or GPC and error for in-season variables. Thus, most models did not arrive at accurate simulations of GY and GPC by accurately simulating preceding growth dynamics. Ensemble simulations, taking either the mean (e-mean) or median (e-median) of simulated values, gave better estimates than any individual model when all variables were considered. Compared to individual models, e-median ranked first in simulating measured GY and third in GPC. The error of e-mean and e-median declined with an increasing number of ensemble members, with little decrease beyond 10 models. We conclude that multimodel ensembles can be used to create new estimators with improved accuracy and consistency in simulating growth dynamics. We argue that these results are applicable to other crop species, and hypothesize that they apply more generally to ecological system models.

model intercomparison

Large Ensemble Exploration of Global Energy Transitions Under National Emissions Pledges

Global climate goals require a transition to a deeply decarbonized energy system. Meeting the objectives of the Paris Agreement through countries' nationally determined contributions and long-term strategies represents a complex problem with consequences across multiple systems shrouded by deep uncertainty. Robust, large-ensemble methods and analyses mapping a wide range of possible future states of the world are needed to help policymakers design effective strategies to meet emissions reduction goals. This study contributes a scenario discovery analysis applied to a large ensemble of 5,760 model realizations generated using the Global Change Analysis Model. Eleven energy-related uncertainties are systematically varied, representing national mitigation pledges, institutional factors, and techno-economic parameters, among others. The resulting ensemble maps how uncertainties impact common energy system metrics used to characterize national and global pathways toward deep decarbonization. Results show globally consistent but regionally variable energy transitions as measured by multiple metrics, including electricity costs and stranded assets. Larger economies and developing regions experience more severe economic outcomes across a broad sampling of uncertainty. The scale of CO 2 removal globally determines how much the energy system can continue to emit, but the relative role of different CO 2 removal options in meeting decarbonization goals varies across regions. Previous studies characterizing uncertainty have typically focused on a few scenarios, and other large-ensemble work has not (to our knowledge) combined this framework with national emissions pledges or institutional factors. Our results underscore the value of large-ensemble scenario discovery for decision support as countries begin to design strategies to meet their goals.

29 ENERGY PLANNING, POLICY, AND ECONOMY

Nanoscale wetting controls reactive Pd ensembles in synthesis of dilute PdAu alloy catalysts

The performance of bimetallic dilute alloy catalysts is largely determined by the size of minority metal ensembles on the nanoparticle surface. By analyzing the synthesis of catalysts comprising Pd 8 Au 92 nanoparticles supported on silica using surface-sensitive techniques, we report that whether Pd overgrowth occurs before or after Au nanoparticle deposition onto the support controls the surface Pd ensemble size and abundance. These differences in Pd ensembles influence catalytic reactivity in H 2 –D 2 isotope exchange and benzaldehyde hydrogenation, which, in correlation with theoretical calculations, is used to elucidate the active site(s) in each reaction. To clarify how the synthetic sequence controls the formation of Pd ensembles, we combine numerical wetting calculations and molecular dynamics simulations (with a machine-learned force field) to visualize Pd deposition and migration on the nanoparticle surface, respectively. Our results suggest that the nanoparticle–support interface restricts nanoparticle accessibility to Pd deposition, which consequently controls the Pd ensemble size, illustrating the critical role of nanoscale wetting phenomena during bimetallic catalyst preparation.

36 MATERIALS SCIENCE

Nonlinear Ensemble Filtering with Diffusion Models: Application to the Surface Quasigeostrophic Dynamics

The intersection between classical data assimilation methods and novel machine learning techniques has attracted significant interest in recent years. Here, we explore another promising solution in which diffusion models are used to formulate a robust nonlinear ensemble filter for sequential data assimilation. Unlike standard machine learning methods, the proposed ensemble score filter (EnSF) is completely training free and can efficiently generate a set of analysis ensemble members. Here, in this study, we apply the EnSF to a surface quasigeostrophic model and compare its performance against the popular local ensemble transform Kalman filter (LETKF), which makes Gaussian assumptions in the analysis step. Numerical tests demonstrate that EnSF maintains stable performance in the absence of localization and for a variety of experimental settings. We find that while LETKF maintains optimal performance in the case of linear observations of the entire state and a perfect model, EnSF shows improvements over LETKF when nonlinear observations are assimilated and the system is subject to unexpected model errors. A spectral decomposition of the analysis results in this nonlinear observation regime shows that the largest improvements over LETKF occur at large scales (small wavenumbers), where LETKF lacks sufficient ensemble spread. Overall, this initial application of EnSF to a geophysical model of intermediate complexity motivates further development of the algorithm for more realistic problems.

Artificial intelligence

Observable-projected ensembles

Measurements in many-body quantum systems can generate non-trivial phenomena, such as preparation of long-range entangled states, dynamical phase transitions, or measurement-altered criticality. Here, we introduce a new measurement scheme that produces an ensemble of mixed states in a subsystem, obtained by measuring a local Hermitian observable on part of its complement. We refer to this as the observable-projected ensemble . Unlike standard projected ensembles-where pure states are generated by projective measurements on the complement-our approach involves projective partial measurements of specific observables. This setup has two main advantages: theoretically, it is amenable to analytical computations, especially within conformal field theories. Experimentally, it requires only a linear number of measurements, rather than an exponential one, to probe the properties of the ensemble. As a first step in exploring the observable-projected ensemble, we investigate its entanglement properties in conformal field theory and perform a detailed analysis of the free compact boson.

Milekhin, Alexey [California Institute of Technolo