Engineering PapersSearch

SEARCH · Engineering Papers

Results for “distributionally robust”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

The CAI Database: 26 Al– 26 Mg Isotope Systematics

We present a publicly available calcium–aluminum-rich inclusion (CAI) database that focuses on the initial 26 Al/ 27 Al 0 ratio in CAIs, designed in a way that researchers in cosmochemistry and astrophysics may find useful. To date, the database contains 497 CAIs from 75 peer-reviewed papers. The CAIs are from all chondrite groups and cover different CAI types, textures, and sizes. The database includes the paper; the host meteorite; the CAI name and type; the 26 Al/ 27 Al 0 , δ 26 Mg$^*_0$, and δ 25 Mg values and their uncertainties; the number of regression points; the maximum 27 Al/ 24 Mg; the mean-squared weighted deviation; the CAI size; and CAI descriptions. We grouped the CAIs in different ways to discuss 26 Al/ 27 Al 0 ratio distributions with implications for the CAI formation timeline. Overall, we agree with previous authors that CAIs have a bimodal 26 Al distribution: CAIs with robust isochrons (n = 151) have a median 26 Al/ 27 Al 0 = 4.8 × 10 −5 (with a 1σ standard error of 0.1), while those with isotopic anomalies (n = 87) have a median 26 Al/ 27 Al 0 = 0.3 × 10 −5 (with a 1σ standard error of 0.2). However, the large standard deviation of both groups (1.3 and 2.3, respectively) indicates that the 26 Al/ 27 Al 0 values scatter significantly within each population. CAI types and groups can have distinct 26 Al/ 27 Al 0 and δ 26 Mg$^*_0$, but the unmelted inclusions (n = 33) have the highest median 26 Al/ 27 Al 0 = 5.1 × 10 −5 and a low median δ 26 Mg$^*_0$ = −0.05‰. We find slightly different 26 Al/ 27 Al 0 distributions between CAI chondrite types, but no differences between petrographic types or sizes. These observations can help us to understand CAI formation in the context of astrophysical models.

Astronomy and AstroPhysics

NanoPSD: A software for automatic detection of Nano-Particle Shape Distribution in electron microscopy images

Accurate quantification of the size and morphology of nanoparticles from electron microscopy (EM) images is essential to understand growth mechanisms, surface reactivity, and functional behavior in nanoscale materials. Manual analysis remains slow, subjective, and difficult to reproduce in large datasets. We introduce NanoPSD (Nano-Particle Shape Distribution), an open-source and fully automated framework for quantitative particle detection and morphology analysis from EM images. NanoPSD integrates adaptive contrast enhancement, polarity-agnostic scale-bar detection, Optical Character Recognition (OCR)-based calibration, and classical segmentation via Otsu thresholding with morphological refinement. Particle contours are used to extract geometric descriptors, including equivalent circular diameter, aspect ratio, circularity, and solidity, enabling automated classification into spherical, rod-like, and aggregate morphologies. The framework supports both single-image and batch processing, generating publication-quality visualizations, LaTeX-ready tables, and structured comma-separated values (CSV) datasets. As a demonstration, we applied NanoPSD to plasma-synthesized nanoparticle samples diagnosed via transmission electron microscopy (TEM). The code produced statistically robust size and morphology distributions spanning a few to tens of nanometers with minimal user supervision. The pipeline demonstrates high reproducibility and scalability, processing large image collections with consistent calibration and output formatting. Its modular design enables seamless integration of future deep-learning-based segmentation models, providing a pathway toward intelligent, data-driven electron microscopy analysis.

36 MATERIALS SCIENCE

Optimization of the FRIB beam dump: a hybrid genetic algorithm and reinforcement learning approach

The operational envelope of high-power-density systems, such as particle accelerators and advanced nuclear energy systems, is critically constrained by the need to manage extreme thermal loads. To address this, we present a novel hybrid optimization framework combining a genetic algorithm (GA) with a soft actor-critic (SAC) deep reinforcement learning agent. This framework was applied to a practical high-heat-flux problem: redesigning the beam dump at the Facility for Rare Isotope Beams (FRIB) for a power upgrade from 20 kW to 50 kW. The resulting design, validated by three-dimensional conjugate heat transfer simulations, suppresses hazardous hot spots and yields a markedly more uniform temperature distribution. This provides a robust operating margin, increasing the average power-handling capability by 72% relative to the current design, demonstrating the framework’s potential to solve complex thermal management challenges in both accelerator technology and advanced nuclear systems.

Accelerator

Confronting Large‐Eddy Simulations With Stereo Camera Data by Means of Reconstructed Hemispheric Cloud Size Distributions

High-resolution hemispheric camera images at a meteorological site in western Germany are used to analyze the multi-dimensional spatial characteristics of continental cumulus cloud fields, and to evaluate Large-Eddy Simulations on this aspect. Traditional non-hemispheric cloud-detecting instruments provide additional reference data. The main model-observation comparison focuses on cloud size distributions (CSDs), employing two methods: (a) directly using three-dimensional model fields, direct CSDs, and (b) using rendered hemispheric images of the model fields as produced by a camera simulator based on path-tracing. In the latter method, both the real and rendered images are used to three-dimensionally reconstruct the cloud fields, yielding hemispheric CSDs. Advantages of hemispheric comparisons over more classic approaches include (a) fair comparisons between model and data, and (b) full use of the enhanced resolutions and hemispheric spatial coverage of the camera imagery. Basic evaluation of the simulations demonstrates good agreement on thermodynamic structure and its diurnal cycle. Cloud heights and cloud cover are intercompared between the model, camera data and other instrumentation, providing insight into their structural differences. A consistent alignment is found between the hemispheric CSDs from both the model and the cameras. Power law fits reveal structurally lower exponents in hemispheric CSDs compared to non-hemispheric CSDs, which particularly caution against directly comparing hemispheric CSDs to non-hemispheric distributions. This result is robust for sample size and fitting method. These findings inform future use of hemispheric camera systems for studying cumulus cloud field morphology and model evaluation.

54 ENVIRONMENTAL SCIENCES

Deep learning-based temporal deconvolution for photon time-of-flight distribution retrieval

The acquisition of the time of flight (ToF) of photons has found numerous applications in the biomedical field. Over the last decades, a few strategies have been proposed to deconvolve the temporal instrument response function (IRF) that distorts the experimental time-resolved data. However, these methods require burdensome computational strategies and regularization terms to mitigate noise contributions. Herein, we propose a deep learning model specifically to perform the deconvolution task in fluorescence lifetime imaging (FLI). The model is trained and validated with representative simulated FLI data with the goal of retrieving the true photon ToF distribution. Its performance and robustness are validated with well-controlled in vitro experiments using three time-resolved imaging modalities with markedly different temporal IRFs. The model aptitude is further established with in vivo preclinical investigation. Overall, these in vitro and in vivo validations demonstrate the flexibility and accuracy of deep learning model-based deconvolution in time-resolved FLI and diffuse optical imaging.

Pandey, Vikas (ORCID:0000000154771095)

A Multi-Sensor Approach for Measuring Bird and Bat Collisions with Offshore Wind Turbines (Final Technical Report)

Collision of birds and bats with wind turbines is a conservation concern for both land-based and offshore wind projects. The fatality rates of birds and bats at land-based turbines are well documented. The measurement strategies on land focus on finding carcasses following collision, estimating the number of carcasses missed through searcher efficiency, carcass persistence trials and carcass fall distributions, and modeling statistically robust fatality rates. Few technologies have been developed to monitor offshore bird and bat collisions, and many that have been developed focused on detecting collisions with large birds. The few studies that have attempted to document collisions at offshore turbines do not account for smaller bodied animals or for collisions that might be missed, which prevents the calculation of statistically robust fatality rates. The overall goal of this report, A Multi-Sensor Approach for Measuring Bird and Bat Collisions with Offshore Wind Turbines (Project), was to develop an effective multi-sensor system for quantifying bird and bat collision rates, specifically for offshore wind facilities. The Project goal and resulting automated collision detection system was achieved through two major technological advancements: 1) refining The Netherlands Organisation for Applied Scientific Research’s (TNO’s) existing WT-Bird® vibration sensing system, that had successfully detected large bird collisions during daytime, to allow for improved detection of smaller birds and bats during both daytime and nighttime hours and 2) improving image processing systems and developing and integrating machine learning algorithms to automatically detect and classify small and large bird and bat collisions with offshore turbines. This final technical report (FTR) summarizes Methods , Results , Conclusions , and Lessons Learned during each of the five Tasks identified for this research and development effort. This FTR includes summaries of the following: Task 1. Initial Engineering Tests to Improve WT-Bird® Task 2. Installation of WT‐Bird® on a Utility-scale Turbine at the National Wind Technology Center – National Renewable Energy Laboratory Task 3. Field Tests and Refinement of the Object Detection System Task 4. Validation of WT-Bird® on a Land-based Turbine Task 5. Preparation for the Implementation of WT-Bird® on an Offshore Turbine. This research and development effort documented successful improvement of the WT Bird® collision detection system to detect small birds and bats, and WT-Bird® is the first collision detection system to validate results compared to land-based post-construction monitoring. The collision trials provide estimates of missed targets that can be used to estimate fatality rates, a significant improvement relative to other offshore collision monitoring systems. Advances were made in developing an edge-processing solution to reduce data storage requirements, which is important if the system is deployed for long periods of time at offshore turbines. The improved WT-Bird® system also provides an important option for wind operators on land or offshore who need to document specific details about when collisions occur, particularly efforts to further research on bat impact minimization, or when standard fatality searches are impractical (e.g. offshore) or inadequate (e.g. challenging locations on land).

17 WIND ENERGY

Resource-Adaptive Federated Text Generation with Differential Privacy

In cross-silo federated learning (FL), sensitive text datasets remain confined to local organizations due to privacy regulations, making repeated training for each downstream task both communication-intensive and privacy-demanding. A promising alternative is to generate differentially private (DP) synthetic datasets that approximate the global distribution and can be reused across tasks. However, pretrained large language models (LLMs) often fail under domain shift, and federated finetuning is hindered by computational heterogeneity: only resource-rich clients can update the model, while weaker clients are excluded, amplifying data skew and the adverse effects of DP noise. We propose a flexible participation framework that adapts to client capacities. Strong clients perform DP federated finetuning, while weak clients contribute through a lightweight DP voting mechanism that refines synthetic text. To ensure the synthetic data mirrors the global dataset, we apply control codes (e.g., labels, topics, metadata) that represent each client’s data proportions and constrain voting to semantically coherent subsets. This two-phase approach requires only a single round of communication for weak clients and integrates contributions from all participants. Experiments show that our framework improves distribution alignment and downstream robustness under DP and heterogeneity.

Wang, Jiayi [ORNL]

Open Set Recognition for Unknown Waveform Classification

This presentation applies open set recognition to classify unknown waveforms, enabling systems to not only identify known types but also reliably detect when waveforms fall outside the training distribution. This approach enhances robustness by avoiding forced misclassification of novel or anomalous signals.

99 - GENERAL AND MISCELLANEOUS

Elliptically-Contoured Tensor-variate Distributions with Application to Image Learning

Statistical analysis of tensor-valued data has largely used the tensor-variate normal (TVN) distribution that may be inadequate for data arising from distributions with heavier or lighter tails. We study a general family of elliptically contoured (EC) TV distributions and derive its characterizations, moments, marginal, and conditional distributions. We describe procedures for maximum likelihood estimation from data that are (1) uncorrelated draws from an EC distribution, (2) from a scale mixture of the TVN distribution, and (3) from an underlying but unknown EC distribution, for which we extend Tyler’s robust estimator. A detailed simulation study highlights the benefits of choosing an EC distribution over the TVN for heavier-tailed data. We develop TV classification rules using discriminant analysis and EC errors and show that they better predict cats and dogs from images in the Animal Faces-HQ dataset than the TVN-based rules. A novel tensor-on-tensor regression and TV analysis of variance (TANOVA) framework under EC errors is also demonstrated to better characterize gender, age, and ethnic origin than the usual TVN-based TANOVA in the celebrated labeled faces of the wild dataset.

97 MATHEMATICS AND COMPUTING

Dark Energy Survey Year 6 Results: Redshift Calibration of the Weak Lensing Source Galaxies

Determining the distribution of redshifts for galaxies in wide-field photometric surveys is essential for robust cosmological studies of weak gravitational lensing. We present the methodology, calibrated redshift distributions, and uncertainties of the final Dark Energy Survey Year 6 (Y6) weak lensing galaxy data, divided into four redshift bins centered at $\langle z \rangle = [0.414, 0.538, 0.846, 1.157]$. We combine independent information from two methods on the full shape of redshift distributions: optical and near-infrared photometry within an improved Self-Organizing Map $p(z)$ (SOMPZ) framework, and cross-correlations with spectroscopic galaxy clustering measurements (WZ), which we demonstrate to be consistent both in terms of the redshift calibration itself and in terms of resulting cosmological constraints within 0.1$σ$. We describe the process used to produce an ensemble of redshift distributions that account for several known sources of uncertainty. Among these, imperfection in the calibration sample due to the lack of faint, representative spectra is the dominant factor. The final uncertainty on mean redshift in each bin is $σ_{\langle z\rangle} = [0.012, 0.008,0.009, 0.024]$. We ensure the robustness of the redshift distributions by leveraging new image simulations and a cross-check with galaxy shape information via the shear ratio (SR) method.

Yin, B. [Duke U.] (ORCID:0009000656049980)

Roadmap and Benchmarking: Privacy in Federated Load Forecasting

Data-driven techniques for energy demand forecasting continue to emerge with promising impacts on distribution grid planning. However, the development of robust and generalizable machine learning models requires that representative high quality training data are available. Distributed energy resources have begun to embed intelligence, gathering large amounts of data on customer demand, behavior, and household devices that are connected to the grid. Though utilities aggregate meter-level demand data for load shaping, demand response, outage management, reliability planning, and billing applications, there lies an inherent privacy concern in sharing consumption data that may identify individual consumer behavioral patterns. Hence, while sharing the data is crucial, the private sensitive customer data must be safeguarded from being exposed or manipulated. In this study, we propose a roadmap for implementing a based privacy preserving framework to support the advancement of data-driven analytics in data-sensitive distributed energy resources environments. The roadmap incorporates federated learning–a distributed training framework, differential privacy–a statistical framework that provides guarantees to safeguard the leakage of sensitive data, secure multiparty computation and homomorphic encryption– techniques for encrypting model gradients and applying secure aggregation on the server. Moreover, we perform baseline experiments on the federated short-term load forecasting (STLF) task using open-source residential load profile datasets, offering insights into the challenges of integrating differential privacy into federated learning.

Abebe, Waqwoya [Oak Ridge National Laboratory (ORN

Single-Phase to Split-Phase Inverters with Advanced Grid Support Functions for Grid-Interactive Applications

This work presents a cost-effective single-phase to split-phase inverter with a reduced switch count, achieving grid interactive performance while maintaining operational efficiency. The proposed system integrates an Andronov-Hopf oscillator based secondary controller, which inherently embeds a nonlinear resistive droop architecture, ensuring rapid dynamic response. A Lyapunov energy function-based primary control enhances transient stability and regulation, while an internal model-based point of common coupling voltage estimation enables cost optimization without additional sensors. Equipped with advanced grid support functionalities, the inverter facilitates seamless distribution system operation with enhanced robustness. The effectiveness of the proposed architecture and control strategy is validated through MATLAB/Simulink and PLECS simulations, demonstrating its feasibility for high-performance grid-supportive applications.

24 POWER TRANSMISSION AND DISTRIBUTION

Constraints on dark photon dark matter from Lyman- α forest simulations and an ultrahigh signal-to-noise quasar spectrum

The ultralight dark photon is a well-motivated, hypothetical dark matter candidate. In a dilute plasma, they can resonantly convert into photons, and heat up the intergalactic medium between galaxies. In this work, we explore the dark photon dark matter parameter space by comparing synthetic Lyman- α forest data from cosmological hydrodynamical simulations to observational data from VLT/UVES of the quasar HE0940-1050 ( z em = 3.09 ). We use a novel flux normalization technique that targets underdense gas, reshaping the flux probability distribution. Not only do we place robust constraints on the kinetic mixing parameter of dark photon dark matter, but notably our findings suggest that this model can still reconcile simulated and observed Doppler parameter distributions of z ∼ 0 Lyman- α lines, as seen by HST/COS. This work opens new pathways for the use of the Lyman- α forest to explore new physics, and can be extended to other scenarios such as primordial black hole evaporation, dark matter decay, and annihilation.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS

Revealing the Hidden Third Dimension of Point Defects in Two-Dimensional MXenes

Point defects govern many important functional properties of two-dimensional (2D) materials. However, resolving the three-dimensional (3D) arrangement of these defects in multi-layer 2D materials remains a fundamental challenge, hindering rational defect engineering. Here, we overcome this limitation using an artificial intelligence-guided electron microscopy workflow to map the 3D topology and clustering of atomic vacancies in Ti3C2TX MXene. Our approach reconstructs the 3D coordinates of vacancies across hundreds of thousands of lattice sites, generating robust statistical insight into their distribution that can be correlated with specific synthesis pathways. This large-scale data enables us to classify a hierarchy of defect structures-from isolated vacancies to nanopores-revealing their preferred formation and interaction mechanisms, as corroborated by molecular dynamics simulations. This work provides a generalizable framework for understanding and ultimately controlling point defects across large volumes, paving the way for the rational design of defect-engineered functional 2D materials.

2D materials

Foundation models for atomistic simulation of chemistry and materials

Conventional computational methods for modeling chemical and materials systems are limited by system size and timescale, forcing a trade-off between quantum-mechanical accuracy and the sampling needed for realistic observables. Large language and vision foundation models — pre-trained on massive datasets using transformer architectures — have revolutionized many fields. It is thus interesting to ask whether a foundation model — subject to suitable data, parameter scaling and training — could enable learned simulations of chemistry and materials. Here, in this study, we review the field of machine-learned interatomic potentials (MLIPs) and posit that scaling up large and diverse chemical and materials datasets and highly expressive architectures using advanced training strategies should result in models that are: more efficient, transferable, robust to out-of-distribution scenarios, and easier to fine-tune to a variety of downstream physical observables than models trained from scratch on small datasets corresponding to specific, targeted atomistic simulation tasks. We provide specific criteria for creating such large-scale MLIP foundation models, coordinated strategies for their development, evaluation and deployment, and highlight potential emergent capabilities that could transform predictive simulations in chemistry and materials science and accelerate discovery across multiple technological domains.

Yuan, Eric C.-Y. [University of California, Berkel

LandScan mosaic enables high-resolution gridded population estimates with explicit uncertainty

Gridded population datasets represent high-resolution distributions of human occupancy, enabling informed decision-making across a broad range of fields. These data products are valuable for assessing environmental risk, urban development, disaster preparedness and resource allocation—areas where accurate population estimates directly enhance policy effectiveness and optimize resource distribution. Despite the importance of gridded population datasets, traditional population modeling approaches often overlook inherent uncertainties in the estimation process. This limitation can create a false sense of certainty in population estimates, potentially leading to flawed decisions by those who rely on the data. To address this methodological gap, we introduce a probabilistic machine learning modeling framework, LandScan Mosaic, that explicitly incorporates uncertainty into the population modeling process. Our approach systematically quantifies uncertainty in three key modeling parameters of the LandScan HD gridded population dataset: building use types, floor counts, and occupancy rates. By employing Monte Carlo simulations, we propagate these uncertainties through the modeling process, yielding probability distributions of population counts in place of deterministic point estimates. We demonstrate the practical application of this framework in Iloilo City, Philippines, using structured decision-making techniques and our probabilistic estimates to identify and prioritize areas most affected by projected flooding, supporting targeted interventions that address both economic and social risks. In doing so, we propose a population-specific approach for incorporating confidence into structured decision making processes. Through a comparative analysis with conventional deterministic approaches and point estimate approaches, including LandScan HD and WorldPop, we evaluate how the incorporation of machine learning and uncertainty influences decision rankings. This research advances population distribution modeling by offering a robust, quantitative approach that explicitly accounts for uncertainty in the underlying data, along with guidance for how users can apply uncertainty in their decision-making.

Environmental sciences

Beyond traditional diagnostics: Identifying active galactic nuclei using spectral energy distribution fitting in DESI data

Active galactic nuclei (AGN) are typically identified through their distinctive X-ray or radio emissions, mid-infrared (MIR) colors, or emission lines. However, each method captures different subsets of AGN due to signal-to-noise (S/N) limitations, redshift coverage, and extinction effects, underscoring the necessity for a multiwavelength approach for comprehensive AGN samples. This study explores the effectiveness of spectral energy distribution (SED) fitting as a robust method for AGN identification. Using CIGALE optical-MIR SED fits on DESI Early Data Release galaxies, we compare SED-based AGN selection (AGNFRAC ≥ 0.1) with traditional methods including BPT diagrams, WISE colors, X-ray, and radio diagnostics. The SED fitting identifies ∼70% of narrow- and broad-line AGN and 87% of WISE-selected AGN. Incorporating high S/N WISE photometry reduces star-forming galaxy contamination from 62% to 15%. Initially, ∼50% of SED-AGN candidates are undetected by standard methods, but additional diagnostics classify ∼85% of these sources, revealing low-ionization nuclear emission-line regions and retired galaxies potentially representing evolved systems with weak AGN activity. Further spectroscopic and multiwavelength analysis will be essential to determine the true AGN nature of these sources. SED fitting provides complementary AGN identification, unifying multiwavelength AGN selections. This approach enables more complete – albeit somewhat contaminated – AGN samples, which are essential for upcoming large-scale surveys where spectroscopic diagnostics may be limited.

Seyfert