Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “aggregation approach”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Aggregation bias and its drivers in large‐scale flood loss estimation: A Massachusetts case study

Abstract Large‐scale estimations of flood losses are often based on spatially aggregated inputs. This makes risk assessments vulnerable to aggregation bias, a well‐studied, sometimes substantial outcome in analyses that model fine‐grained spatial phenomena at coarse spatial units. To evaluate this potential in the context of large‐scale flood risk assessments, we use data from a high‐resolution flood hazard model and structure inventory for over 1.3 million properties in Massachusetts and examine how prominent data aggregation approaches affect the magnitude and spatial distribution of flood loss estimates. All considered aggregation approaches rely on aggregate structure inventories but differ in whether flood hazard is also aggregated. We find that aggregating only structure inventories slightly underestimates overall losses (−10% bias), and when flood hazard data is spatially aggregated to even relatively small spatial units (census block), statewide aggregation bias can reach +366%. All aggregation‐based procedures fail to capture the spatial covariation of inputs distributions in the upper tails that disproportionately generate total expected losses. Our findings are robust to several key assumptions, add important context to published risk assessments and highlight opportunities to improve flood loss estimation uncertainty quantification.

54 ENVIRONMENTAL SCIENCES↗

AssessCCUS: An Integrated Approach for Aggregating Resources to Enable Techno-Economic and Life Cycle Assessment of Carbon Management Technologies

Carbon capture, utilization, and storage (CCUS) - also sometimes known as carbon management - technologies are becoming an increasingly important part of the portfolio of technologies necessary to mitigate climate change and defossilize industrial production systems (Sick, 2021). These technologies capture carbon dioxide from industrial point sources or from the atmosphere directly and then either sequester it or use it as a carbon source in valuable products. Potential utilization pathways include, but are not limited to, concrete, fuels, and certain commodity chemicals, and sequestration pathways can include permanent geological storage or temporary storage in natural sinks ranging from forests to agricultural soil. Regardless of the pathway, assessment of the economic and environmental performance of the technologies is important for understanding their potential scalability and impact as well as developing plans to minimize life cycle costs and potential environmental trade-offs. A full discussion of potential trade-offs associated with CCUS is outside the scope of this article, but promoting assessment broadly helps to stimulate important conversations about the benefits and drawbacks of any particular technological choice.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Traceable and Scalable Food Balance Sheets from Agricultural Commodity Supply and Utilization Accounts (2010–2022)

Abstract The Food Balance Sheets (FBS), compiled by the Food and Agriculture Organization (FAO), serve as a cornerstone dataset for studies on agricultural development, food security, and dietary health, providing a broad overview of global and regional food systems. However, its limited transparency and scalability hinder its application in empirical analysis and multisector dynamic modeling. Here, we present a traceable Food Balance Sheets (T-FBS) dataset, developed from detailed Supply Utilization Accounts (SUA) using a novel Primary Commodity equivalent (PCe) aggregation approach. This framework enables the aggregation of commodity flows along supply chains while ensuring consistency and balance across multiple dimensions. The T-FBS dataset includes 57 PCe commodities across 195 regions for the period 2010–2022, consolidated from over 500 SUA products. While T-FBS closely aligns with FAO-FBS at aggregate levels for dietary energy and macronutrients, it identifies key uncertainties in other elements (e.g., feed, trade, stocks). By enhancing methodological transparency, traceability, and scalability, T-FBS strengthens the robustness of food system studies and fosters future research and collaboration within the open-source community.

agriculture↗

Multi-Kernel Adaptive Support Vector Machine for Scalable Predictive Maintenance

Application of data-driven solutions across an industry is challenging, since the data are often stored locally, and increasing privacy and security concerns restrict access to the data. In addition, it is highly unlikely that all potential data patterns are captured in a single data source. Because it is highly unlikely that all potential data patterns are captured in a single data source, machine learning (ML) models developed from a single source cannot be robust enough. An alternative is to train the ML model at each source and develop a distributed knowledge discovery and aggregation approach to build global knowledge. In this paper, we develop and demonstrate a distributed ML model, federated transfer learning (FTL), using a multi-kernel-based adaptive support vector machine (MK-A-SVM). For federated learning (FL), the multi-kernel (MK) approach enables feature-specific model aggregation under data heterogeneity; whereas for transfer learning (TL) the adaptive model enables utilization of an aggregated model from a different task. The proposed approach is validated using nuclear power plant (NPP) vertical motor-driven pump data to predict the health condition of vertical motor-driven pumps as an anomaly detection. The efficiency of the proposed approach is also quantified and compared with neural network.

42 ENGINEERING↗

What Technical Choices Matter to Characterize Heat Wave and Cold Snap Events in Support of Bulk Power Grid Reliability Studies?

Extreme weather events, such as Heat Waves (HW) and Cold Snaps (CS), pose significant risks to the power grid. The United States (U.S.) Federal Energy Regulatory Commission Order No. 896 mandates regional coordination standards that account for extreme thermal events. However, the lack of a universal definition for extreme thermal events may lead to inconsistent compliance efforts among neighboring entities, undermining the reliability of the transmission system. This study directly addresses this challenge by systematically evaluating how varying technical choices in defining HW and CS fundamentally impact the characterization and ranking of extreme events for power grid reliability studies. We used 12 event definitions and multiple temperature spatial aggregation approaches to construct historical (1980–2024) regional extreme thermal event libraries across North American Electric Reliability Corporation (NERC) subregions in the conterminous U.S. We examined the sensitivity of event characteristics (e.g., duration, frequency, intensity, and spatial coverage) to different definitions. While some definitions produced similar libraries and top event rankings, definitions based on moving-window-averaged temperatures yielded markedly different characteristics. Spatial aggregation methods had minimal impact on heat wave or cold snap intensity, frequency and duration but significantly influenced spatial coverage. The top events identified across different aggregation methods were consistent, but their ranking order varied. These findings offer critical insights for characterizing and selecting extreme thermal events and for supporting local and cross-regional coordination as required by reliability standards.

Wan, Heng [Pacific Northwest National Laboratory (↗

Resolving Configurational Disorder for Impurities in a Low-Entropy Phase

Hematite (α-Fe 2 O 3 ) exerts a strong control over the transport of minor but critical metals in the environment and is used in multiple industrial applications; the photocatalysis community has explored the properties of hematite nanoparticles over a wide range of transition metal dopants. Nonetheless, simplistic assumptions are used to rationalize the local coordination environment of impurities in hematite. Here, we use ab initio molecular dynamics (AIMD)-guided structural analysis to model the extended X-ray absorption fine structure (EXAFS) of Cu 2+ - and Zn 2+ -doped hematite nanoparticles. Specific defect–impurity associations were identified, and the local coordination environments of Cu and Zn both displayed considerable configurational disorder that, in aggregate, approached Jahn–Teller-like distortion for Cu but, in contrast, maintained hematite-like symmetry for Zn. This study highlights the role of defects in accommodating impurities in a nominally low-entropy phase and the limits to traditional shell-by-shell fitting of EXAFS for dopants/impurities in unprecedented bonding environments.

36 MATERIALS SCIENCE↗

Automated Quantification of Wind Turbine Blade Leading Edge Erosion from Field Images

Wind turbine blade leading edge erosion is a major source of power production loss and early detection benefits optimization of repair strategies. Two machine learning (ML) models are developed and evaluated for automated quantification of the areal extent, morphology and nature (deep, shallow) of damage from field images. The supervised ML model employs convolutional neural networks (CNN) and learns features (specific types of damage) present in an annotated set of training images. The unsupervised approach aggregates pixel intensity thresholding with calculation of pixel-by-pixel shadow ratio (PTS) to independently identify features within images. The models are developed and tested using a dataset of 140 field images. The images sample across a range of blade orientation, aspect ratio, lighting and resolution. Each model (CNN v PTS) is applied to quantify the percent area of the visible blade that is damaged and classifies the damage into deep or shallow using only the images as input. Both models successfully identify approximately 65% of total damage area in the independent images, and both perform better at quantifying deep damage. The CNN is more successful at identifying shallow damage and exhibits better performance when applied to the images after they are preprocessed to a common blade orientation.

Aird, Jeanie A.↗

Stream Temperature Predictions for River Basin Management in the Pacific Northwest and Mid-Atlantic Regions Using Machine Learning

Stream temperature (Ts) is an important water quality parameter that affects ecosystem health and human water use for beneficial purposes. Accurate Ts predictions at different spatial and temporal scales can inform water management decisions that account for the effects of changing climate and extreme events. In particular, widespread predictions of Ts in unmonitored stream reaches can enable decision makers to be responsive to changes caused by unforeseen disturbances. In this study, we demonstrate the use of classical machine learning (ML) models, support vector regression and gradient boosted trees (XGBoost), for monthly Ts predictions in 78 pristine and human-impacted catchments of the Mid-Atlantic and Pacific Northwest hydrologic regions spanning different geologies, climate, and land use. The ML models were trained using long-term monitoring data from 1980–2020 for three scenarios: (1) temporal predictions at a single site, (2) temporal predictions for multiple sites within a region, and (3) spatiotemporal predictions in unmonitored basins (PUB). In the first two scenarios, the ML models predicted Ts with median root mean squared errors (RMSE) of 0.69–0.84 °C and 0.92–1.02 °C across different model types for the temporal predictions at single and multiple sites respectively. For the PUB scenario, we used a bootstrap aggregation approach using models trained with different subsets of data, for which an ensemble XGBoost implementation outperformed all other modeling configurations (median RMSE 0.62 °C).The ML models improved median monthly Ts estimates compared to baseline statistical multi-linear regression models by 15–48% depending on the site and scenario. Air temperature was found to be the primary driver of monthly Ts for all sites, with secondary influence of month of the year (seasonality) and solar radiation, while discharge was a significant predictor at only 10 sites. The predictive performance of the ML models was robust to configuration changes in model setup and inputs, but was influenced by the distance to the nearest dam with RMSE <1 °C at sites situated greater than 16 and 44 km from a dam for the temporal single site and regional scenarios, and over 1.4 km from a dam for the PUB scenario. Our results show that classical ML models with solely meteorological inputs can be used for spatial and temporal predictions of monthly Ts in pristine and managed basins with reasonable (<1 °C) accuracy for most locations.

54 ENVIRONMENTAL SCIENCES↗

Latent Neural ODE for Integrating Multi-Timescale Measurements in Smart Distribution Grids

Under a smart grid paradigm, there has been an increase in sensor installations to enhance situational awareness. The measurements from these sensors can be leveraged for real-time monitoring, control, and protection. However, these measurements are typically irregularly sampled. These measure-ments may also be intermittent due to communication bandwidth limitations. To tackle this problem, this paper proposes a novel latent neural ordinary differential equations (LODE) approach to aggregate the unevenly sampled multivariate time-series measurements. The proposed approach is flexible in performing both imputations and predictions while being computationally efficient. Simulation results on IEEE 37 bus test systems illustrate the efficiency of the proposed approach.

multi time-scale measurements↗

Background-Aware 3-D Point Cloud Segmentation With Dynamic Point Feature Aggregation

With the proliferation of LiDAR sensors and 3-D vision cameras, 3-D point cloud analysis has attracted significant attention in recent years. In this article, we propose a novel 3-D point cloud learning network, referred to as dynamic point feature aggregation network (DPFA-Net), by selectively performing the neighborhood feature aggregation (FA) with dynamic pooling and an attention mechanism. DPFA-Net has two variants for semantic segmentation and classification of 3-D point clouds. As the core module of the DPFA-Net, we propose an FA layer, in which features of the dynamic neighborhood of each point are aggregated via a self-attention mechanism. In contrast to other segmentation models, which aggregate features from fixed neighborhoods, our approach can aggregate features from different neighbors in different layers providing a more selective and broader view to the query points and focusing more on the relevant features in a local neighborhood. In addition, to further improve the performance of semantic segmentation, we exploit the background–foreground (BF) information and present two novel approaches, namely, two-stage BF-Net and BF regularization. Experimental results show that the proposed DPFA-Net achieves the state-of-the-art overall accuracy score of 89.22% for semantic segmentation on the Stanford large-scale 3-D Indoor Spaces (S3DIS) dataset and provides consistently satisfactory performance across different tasks of semantic segmentation, part segmentation, and 3-D object classification. Furthermore, our model achieves 93.1% accuracy on the ModelNet40 dataset and provides a mean shape intersection-over-union (IoU) value of 85.5% for part segmentation on the ShapeNet-Part dataset. It is a also computationally more efficient compared to other methods.

3-D↗

An interregional optimization approach for time series aggregation in continent-scale electricity system models

Modeling electric power systems with high shares of weather-dependent resources requires tradeoffs between temporal, spatial, and operational resolution. Many studies perform time series aggregation using clustering algorithms to reduce the temporal dimension, but when modeling continent-scale electricity systems that are large enough to contain multiple independent weather systems, this approach requires large numbers of representative periods to minimize errors in regional wind and solar capacity factors. Here, a new optimization-based approach for representative period selection and weighting is introduced that minimizes regional errors in average renewable capacity factors and electricity demand. The method delivers higher regional fidelity with fewer representative periods than alternative clustering methods when applied to wind, solar, and demand profiles for the contiguous United States. When representative periods are selected from multiple weather years, the optimized method reproduces regional averages with lower error than a complete 365-day time series from any single weather year. The method identifies only representative (as opposed to outlying) periods but can be combined with an iterative "stress period" identification approach to guide efficient decision-making considering both average and high-risk weather conditions.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Aggregation of Inverter-Based Resources for Modeling and Simulation

In order to conduct system dynamic studies, it is necessary to have dynamic models of both inverter and plant levels. Detailed and aggregated modeling approaches are two essential options. The detailed modeling method involves capturing the dynamic characteristics of each individual device (e.g., wind turbine or PV array), as well as their interconnections. However, as the scale of the IBR plant increases, the complexity and computation time required for detailed modeling also increase. On the other hand, aggregated modeling offers a more efficient way of representing large-scale IBRs in power system dynamic studies. This approach involves aggregating a large number of wind turbines, PV arrays, inverters, and/or plant controllers into one or a smaller number of equivalent models. In order to analyze the impact of a high-level IBR penetration in power systems, it is important to develop accurate and computationally efficient models for both the detailed and aggregated methods.

14 SOLAR ENERGY↗

Stochastic Approaches for Calculating and Aggregating Detection Probabilities for Nuclear Material Diversion

The authors built and tested a stochastic simulation to estimate achieved detection probabilities (DPs) on a stratum basis, over a tailorable range of diverted amounts from 0 to 2 significant quantities (SQ), using typical International Atomic Energy Agency (IAEA) inspection data: i.e., SQ in stratum, number of items, number of gross/partial/bias defect measurements conducted, and realistic relative standard deviation (RSD) values for typical IAEA verification measurements. For bulk strata, the model calculates achieved DP at 0.01 SQ diversion increments; for item strata, the model calculates DP using the smallest realistic diversion increment (e.g., a plate, pin, or coupon). After successfully benchmarking against IAEA deterministic models, the simulation was used to test the sensitivity of DP to certain standard assumptions and selected input parameters. First, the equal defect assumption was tested; the results suggest significant complexity in the effectiveness of partial defect measurements. Next, the authors explored the sensitivity of DP to the assumed RSD of attribute tests. Then, the authors compared non-normal models for instrument performance (e.g., logistic, step, or arbitrary functions) to the typical results from a normal distribution (characterized by RSD). This last comparison was supplemented with experimentally derived performance data for an HM-5 gamma spectrometer. The HM-5 was used to make enrichment measurements on both LEU and HEU MTR fuel elements as plates were removed, and the results fit with logistic and step curves and applied in the simulation. These stochastic DP results were compared to DP estimates from a deterministic model assuming a normal curve and typical RSD, yielding insights that could improve effectiveness in the field. These early results illustrate the potential of stochastic models to better understand achieved DP and to improve safeguards effectiveness.

98 NUCLEAR DISARMAMENT, SAFEGUARDS, AND PHYSICAL P↗

Aggregating in vitro-grown adipocytes to produce macroscale cell-cultured fat tissue with tunable lipid compositions for food applications

We present a method of producing bulk cell-cultured fat tissue for food applications. Mass transport limitations (nutrients, oxygen, waste diffusion) of macroscale 3D tissue culture are circumvented by initially culturing murine or porcine adipocytes in 2D, after which bulk fat tissue is produced by mechanically harvesting and aggregating the lipid-filled adipocytes into 3D constructs using alginate or transglutaminase binders. The 3D fat tissues were visually similar to fat tissue harvested from animals, with matching textures based on uniaxial compression tests. The mechanical properties of cultured fat tissues were based on binder choice and concentration, and changes in the fatty acid compositions of cellular triacylglyceride and phospholipids were observed after lipid supplementation (soybean oil) during in vitro culture. This approach of aggregating individual adipocytes into a bulk 3D tissue provides a scalable and versatile strategy to produce cultured fat tissue for food-related applications, thereby addressing a key obstacle in cultivated meat production.

Yuen Jr, John Se Kit (ORCID:0000000198541654)↗

Trustworthiness modeling and evaluation for a nearly autonomous management and control system

The Nearly Autonomous Management and Control (NAMAC) system supports the advanced reactor operation by recommending control actions to operators based on real-time measurements and digital twins (DTs) learning from the knowledge base. To enable the safe and reliable use of autonomous technologies, NAMAC and its recommendations should be trustworthy to operators and regulators at both the design and operation stages. This study proposes a NAMAC trustworthiness modeling and evaluation framework supported by trustworthiness ontologies and evidence-based approaches. The development-time and run-time ontologies are separately constructed and then converted to Bayesian networks to quantitatively evaluate the NAMAC trustworthiness. This evaluation is demonstrated by collecting and characterizing evidence from NAMAC practices, such as the development and assessment of the NAMAC system, data coverage assessment, and the training and optimizations of neural-network-based DTs. Our proposed approach can aggregate various trustworthiness attributes of complex artificial-intelligence-supported systems for safety-critical applications. It also considers the interaction between different DTs and extends beyond the trustworthiness evaluation of a single DT. In conclusion, the evidence-based method enhances the transparency of the trustworthiness modeling and evaluation processes and helps identify uncertainties and subjectivity involved in the processes.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

When to use rsync

We have endeavored to show, using a series of data transfer results obtained from two testbeds, when to use the popular data copying tool rsync and related tools. Tests have been conducted in local area network (LAN) and wide area network (WAN) environments. We conclude that for files in a certain size range and network latency &lE; 10 ms round trip time (RTT), rsync is still useful for data moving tasks in the category 4 of the U.S. DOE Technical Report “Data Movement Categories”. For more demanding data movement requirements, tools of different classes are suggested. Sample histograms from two DOE user facilities are provided to further support our conclusions.

97 MATHEMATICS AND COMPUTING↗