Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “analysis and statistical methods”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

U.S. Average Nitrogen Fertilizer Production Baseline Documentation For 45Q Life Cycle Analysis: Version 1.0 (2018-2022)

U.S. Average Benchmark life cycle inventory documentation for nitrogenous fertilizer production for calendar years 2018 - 2022. The documentation includes a technical report describing data sources and statistical methods used and an Excel file showing calculations. The calculations performed are used in the 45Q benchmark json-ld dataset, compatible with the NETL CO2U LCA Guidance Toolkit. Summary impact results will be included in the NETL CO2U LCA Documentation Spreadsheet within the Toolkit. https://www.netl.doe.gov/energy-analysis/details?id=7c0db05a-231e-40a5-b5fd-112a7c001834

09 BIOMASS FUELS↗

Statistical data analysis of x-ray spectroscopy data enabled by neural network accelerated Bayesian inference

Bayesian inference applied to x-ray spectroscopy data analysis enables uncertainty quantification necessary to rigorously test theoretical models. However, when comparing to data, detailed atomic physics and radiation transfer calculations of x-ray emission from non-uniform plasma conditions are typically too slow to be performed in line with statistical sampling methods, such as Markov Chain Monte Carlo sampling. Furthermore, differences in transition energies and x-ray opacities often make direct comparisons between simulated and measured spectra unreliable. Here, we present a spectral decomposition method that allows for corrections to line positions and bound–bound opacities to best fit experimental data, with the goal of providing quantitative feedback to improve the underlying theoretical models and guide future experiments. In this work, we use a neural network (NN) surrogate model to replace spectral calculations of isobaric hot-spots created in Kr-doped implosions at the National Ignition Facility. The NN was trained on calculations of x-ray spectra using an isobaric hot-spot model post-processed with Cretin, a multi-species atomic kinetics and radiation code. The speedup provided by the NN model to generate x-ray emission spectra enables statistical analysis of parameterized models with sufficient detail to accurately represent the physical system and extract the plasma parameters of interest.

47 OTHER INSTRUMENTATION↗

Unique & challenging aspects of plutonium metal standards exchange program for actinide measurements

The Los Alamos National Laboratory exchange program is the only program of its kind for the distribution of plutonium (Pu) standards materials with a range of impurity contents to multiple laboratories for destructive measurements of elemental concentration. This paper discusses statistical methods used to address challenges in Pu metal exchange data by way of two case studies. Challenges include how to evaluate a data set when a large fraction of the values are minimum detection limits (MDLs), and how to determine potential outliers with limited in-formation on the true spread of the data.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Modal Field Reconstruction in Resonant Cavities in the Fundamental and Undermoded Frequency Regimes

Theory, simulations, and experiments are presented that demonstrate reconstruction of electromagnetic fields in a cavity from sparse probe measurements. Such techniques are often referred to as virtual sensing, allowing fields at unobserved locations to be predicted. These methods are appropriate for the fundamental and undermoded regimes, providing the ability to estimate fields (and shielding effectiveness) throughout an arbitrarily shaped cavity from a few judiciously spaced probes. A modal simulation method is implemented that allows the response of arbitrarily shaped cavities to be rapidly computed with respect to varying probe locations and slot parameters, enabling statistical analysis of probe placement on reconstruction performance. A cylindrical vessel with numerous probe holes is developed for experiments, referred to as Perforated Vessel 2 (PV2). Experiments are performed on the vessel with and without a steel box inside, where transmit power is delivered into the vessel either through probes (probe injection) or through slots using an external antenna (slot excitation). Simulations and experiments illustrate that when the number of probes is minimal (equal to the number of mode coefficients to be estimated at each frequency), probe placement is critical to avoid missed peaks and to have acceptable reconstruction error. Probe placement becomes less important as the number of probes is increased, but care is still required to avoid probe locations giving poor performance.

42 ENGINEERING↗

Mining Product Reviews for Important Product Features of Refurbished iPhones

Problem: Remanufacturers want to increase consumer interest in refurbished products, which motivates the need to understand which product features are important to buyers of refurbished products such as mobile phones. Research Questions: This study addresses two questions. First, which product features are most important for buyers of refurbished iPhones? Second, how do those preferences differ from the preferences of buyers of new iPhones? Methods: Online reviews of iPhones are obtained and converted into a document–term matrix. Using this text model, three subsets of features are identified using statistical analysis of frequency of mention: most frequent, average, and least frequent. A logistic regression (LR) model is then used to identify which features are most predictive of whether a review is for a new or refurbished phone. Results: Buyers of refurbished phones mention battery health, screen/display, shell condition, and brand significantly more often than other features. Directly contrasting reviews of refurbished versus new phones shows that shell condition, brand, speaker, and charger are found to be the most predictive product features indicated in reviews for refurbished phones. Of those, the shell condition is significantly more predictive than the others. Implications: The results identify product features that remanufacturers of iPhones can emphasize to increase customer demand.

Anisi, Atefeh↗

Raptor

Raptor is an efficient Python-based tool for predicting the formation and morphology of stochastic lack of fusion defects in metal AM processes. A major obstacle for the qualification and certification of additively manufactured parts in critical applications continues to be performance variability caused in part by porosity-related defects. High-fidelity process models that could predict these defect features are currently too computationally expensive for component-level analysis. To address this, Raptor employs a high-performance geometric method to model the dynamic melt pool rather than relying on computationally intensive thermal fluid dynamics. This allows Raptor to rapidly identify regions of unmelted material that correspond to lack of fusion pores. The efficiency of this approach significantly reduces the time and resources needed for generating 3D defect predictions, which enables users to conduct large-scale parameter studies and evaluate how process variations affect part quality. The framework offers operational flexibility; users can execute simulations through a simple command line interface or integrate core functions as a library within larger computational workflows. Simulation outputs include 3D porosity maps for visualization and tools for quantitative morphological analysis. These results are suitable for direct comparison with experimental characterization data from methods such as X-ray computed tomography and can be used for statistical process optimization.

Subraveti, Vamsi [Vanderbilt Univ., Nashville, TN ↗

Highly accelerated life testing (HALT): A review from a statistical perspective

Despite its use in one form or another for at least four decades, HALT and related techniques [e.g., highly accelerated-stress screening (HASS) and stress audits (HASA)] are not well understood within the statistical community and remain controversial. This largely reflects a conflict in motivation between engineers, testing under harsh conditions to discover and eliminate failure modes, and statisticians, taking a more cautious approach to develop quantitative estimates of parameters such as mean time between failures (MTBF). Here, this review article will clarify HALT concepts and methods and explain where it fits within the universe of methods that involve the application of accelerating factors to compress the time required to evaluate or enhance product reliability. A major distinction is between methods such as HALT, a high-stress test-analyze-fix-test iterative process directed at improving reliability by discovering and fixing weak points in a design, and quantitative accelerated life testing (QALT), whose goal is the estimation of product life for a fixed design. We discuss methods such as physics of failure that offer some hope of bridging the gap between the qualitative nature of HALT, and purely quantitative statistical methods. We present a variety of engineering applications of HALT including metal fatigue, piping and pressure vessels, structural damage, radiation damage, and rotating machinery. We also discuss potential synergies between HALT and QALT, such as rapid identification, through HALT, of failure modes requiring quantitative analysis. For further study, extensive references to the applicable literature are provided as well as an appendix that describes related methods.

97 MATHEMATICS AND COMPUTING↗

U.S. Average Market Carbon Dioxide Production Baseline Documentation For 45Q Life Cycle Analysis: Version 1.0 (2018-2022)

U.S. Average Benchmark life cycle inventory documentation for market carbon dioxide production for calendar years 2018 - 2022. The documentation includes a technical report describing data sources and statistical methods used and an Excel file showing calculations. The calculations performed are used in the 45Q benchmark json-ld dataset, compatible with the NETL CO2U LCA Guidance Toolkit. Summary impact results will be included in the NETL CO2U LCA Documentation Spreadsheet within the Toolkit. To access the model referenced in the report, please visit https://www.netl.doe.gov/energy-analysis/details?id=21915ba8-d2cf-43ea-b9b2-1d76363a46a7

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Geospatial analysis of preterm and small-for-gestational age births in Washington D.C.

Background: This study is based on the recognition that adverse pregnancy outcomes significantly affect maternal and infant health, leading to increased morbidity and mortality. These outcomes are shaped by a complex interplay of individual-level factors—like maternal age and education—and community-level influences, including socio-economic status and access to healthcare. Understanding these determinants is crucial for developing effective public health strategies, especially for marginalized populations, by identifying high-risk areas and informing targeted interventions that address both individual and structural barriers. Methods: We utilized geospatial analysis to explore the association between individual- and community-level factors and adverse pregnancy outcomes, specifically preterm birth (PTB) and small-for-gestational-age (SGA) birthweight in Washington, D.C. We used Empirical Bayes smoothing methods to calculate rates of adverse birth outcomes from 2010 to 2018 at the U.S. Census tract–level. Spatial scan statistics were used to investigate if adverse birth outcomes clustered in specific areas. ANOVA tests were conducted for individual- and community-level factors within identified clusters. Results: Spatial analysis identified significant high-risk clusters for PTB and SGA infants primarily in southeastern Washington, D.C., particularly in Wards 7 and 8. Individuals residing within these clusters experienced a 47% increased risk of PTB (RR = 1.467) and a 56% increased risk of SGA (RR = 1.560) compared to those outside clusters. Space–time analysis revealed temporal variation, with PTB clusters persisting from 2011 to 2014 and SGA clusters extending through 2017. Compared to low-risk clusters, high-risk clusters had younger birthing individuals (mean age ~26.5 vs. ~33 years), lower maternal college degree attainment (~20% vs. ~80%), higher rates of late or no prenatal care (~16% vs. 11%), and increased prevalence of smoking and hypertension (all P < 0.001). Community-level indicators showed lower median household incomes ($\$40,000$ vs. ~$\$105,000$), greater poverty (~16% vs. ~7% below $\$10,000$/year), higher public assistance use (~32% vs. ~5%), and reduced healthcare access (greater distances to emergency and specialty care) in high-risk areas (all P < 0.001). Neighborhood deprivation indices were significantly elevated, commutes were longer, and population density was lower in these clusters. These findings highlight that adverse birth outcomes cluster in neighborhoods with pronounced socioeconomic and health disparities. Conclusion: High-risk birth clusters highlight intertwined factors: individual, socio-economic, and geographic. Addressing these requires comprehensive interventions focusing on social and structural determinants of health.

Birth outcomes↗

A Computational Review of Privacy-Preserving Mechanisms for the Smart Grid

Smart grid technologies have rapidly become one of the largest and most comprehensive sources of data for the modern utility. For the most part, data streams are seen as an essential tool that enable utilities to carry their day-to-day business operations, but they also create the need for efficient and secure data management strategies. In the context of the smart grid, ensuring data privacy is becoming an increasing concern due to a combination of factors that range from shifts in operational paradigms and rapid technology evolution to changes in legislation. Furthermore, researchers have highlighted the risks associated with improperly protected energy records. For example, energy consumption data from homes could be used to infer the behaviors and habits of home occupants through activity recognition or user profiling (Fan, 2017), which may lead to unfair service pricing, targeted advertising, or other personal security violations. Similarly, Electric Vehicles’ (EVs) charging metadata could be used to reveal private information about the owner such as their payment methods, preferred charging stations, and other locational and timing information that could be used to reconstruct the vehicle owner’s behaviors. The privacy of user data, even when used for statistical analysis or machine learning training processes, also needs to be carefully considered, as an individual’s private traits may still be vulnerable if their inclusion/exclusion greatly impacts the result or could be linked to a public dataset through cross-reference. The breach of user privacy also has severe impacts for organizations that store, transmit, or work on the data in the form of diminishing the public’s trust in them while potentially incurring legal consequences (e.g., fines and suspensions under the European Union General Data Protection Regulation, Health Insurance Portability and Accountability Act, etc.). Because of these risks, several privacy-preserving mechanisms are available to help organizations comply with privacy legislations and prevent the unauthorized and malicious use of user data. In light of these concerns, this report focuses on performing a computational review of privacy-preserving mechanisms that have received a significant amount of interest in literature. It specifically focuses on 1) homomorphic encryption, 2) zero-knowledge proofs, 3) differential privacy, and 4) federated learning. It is worth noting that although many of the methods presented in this document rely on cryptographic primitives, their intent is not to provide perfect secrecy, but rather to enable users to maintain privacy, and thus they shall not be compared or equated to other constructs that are aimed to address cybersecurity constructs.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Opportunities in AI/ML for the Rubin LSST Dark Energy Science Collaboration

The Vera C. Rubin Observatory's Legacy Survey of Space and Time (LSST) will produce unprecedented volumes of heterogeneous astronomical data (images, catalogs, and alerts) that challenge traditional analysis pipelines. The LSST Dark Energy Science Collaboration (DESC) aims to derive robust constraints on dark energy and dark matter from these data, requiring methods that are statistically powerful, scalable, and operationally reliable. Artificial intelligence and machine learning (AI/ML) are already embedded across DESC science workflows, from photometric redshifts and transient classification to weak lensing inference and cosmological simulations. Yet their utility for precision cosmology hinges on trustworthy uncertainty quantification, robustness to covariate shift and model misspecification, and reproducible integration within scientific pipelines. This white paper surveys the current landscape of AI/ML across DESC's primary cosmological probes and cross-cutting analyses, revealing that the same core methodologies and fundamental challenges recur across disparate science cases. Since progress on these cross-cutting challenges would benefit multiple probes simultaneously, we identify key methodological research priorities, including Bayesian inference at scale, physics-informed methods, validation frameworks, and active learning for discovery. With an eye on emerging techniques, we also explore the potential of the latest foundation model methodologies and LLM-driven agentic AI systems to reshape DESC workflows, provided their deployment is coupled with rigorous evaluation and governance. Finally, we discuss critical software, computing, data infrastructure, and human capital requirements for the successful deployment of these new methodologies, and consider associated risks and opportunities for broader coordination with external actors.

Aubourg, Eric [APC, Paris] (ORCID:000000025592023X↗

elm-diagnostics

elm-diagnostics is a Python package for computing diagnostic analyses and visualizations for the E3SM Land Model (ELM) component and is meant to support new feature development in ELM. The tool reads model history files and performs quantitative analyses including budget-closure checking, variable transformations, temporal aggregations, and statistical summaries to support model evaluation, validation, and scientific interpretation. The framework is designed for extensibility, with modular architecture enabling straightforward addition of new diagnostic methods, derived variables, analysis types, visualization approaches, and model-specific adaptations

Hoffman, Matt [Los Alamos National Laboratory]↗

Measuring labor productivity dynamics in U.S. industrial and electric power sectors: a case study (2014–2023)

This study proposes a subsystem methodology for measuring labor productivity in the U.S. industrial and electric power sectors by leveraging public data available between 2014 and 2023. Building on Pasinetti’s framework and subsequent developments, the approach employs Vertically Integrated Sectors (VIS) to account for both direct and indirect productivity effects. The novelty of this work is twofold. First, it enables the estimation of productivity trends over time, providing a robust foundation for empirical analysis. Second, it applies the methodology to a case study of the electric generation sector, highlighting its practical relevance. Using data from the Bureau of Economic Analysis, the Bureau of Labor Statistics, and the Impact Analysis for Planning (IMPLAN) tool, the study reveals significant discrepancies between conventional productivity measures and those derived from the VIS approach. Furthermore, the proposed method aligns with the principles of Integrated Energy Systems by capturing the interrelations among generation, distribution, storage, and consumption. This alignment underscores its utility and applications for energy-related policy and planning. Overall, the findings contribute to more precise labor productivity assessments, supporting informed decision-making and future research. Additionally, the method highlights the importance of considering the whole supply chain, providing interrelated metrics for labor productivity which includes both, direct and indirect effects on the final labor productivity metric. By incorporating intersectoral dependencies, this method offers a more comprehensive and accurate measure of labor productivity compared to traditional metrics.

11 - NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Machine learning insights into microstructural origins of transport and mechanical properties in porous microstructures

Multifunctional porous materials are increasingly needed across various fields, but their complex microstructures create significant challenges due to the intricate microstructure-property relationships. This complexity, combined with limitations of traditional analysis methods, hinders efforts to understand and optimize microstructure–property relationships. Here, to address this, we integrate physics-based mesoscale modeling with interpretable machine learning (ML) to uncover how microstructural features govern effective diffusivity and elastic modulus. At constant porosity, we show diffusivity varies by over 150 × and modulus by ∼50 ×, highlighting the power of microstructure engineering. Statistical analysis reveals bimodal behavior in diffusivity and unimodal in modulus. ML identifies connectivity as the dominant factor, while modulus is also sensitive to domain size and feature interactions. Controlled simulations further highlight domain shape as a critical feature for modulus. This framework enables efficient exploration of microstructure-property correlations, offering new insights to guide the design of advanced porous materials.

Bicontinuous microstructure↗

Real-time neutron multiplicity and source localization for criticality safety during fuel debris removal

Advancing neutron detection and analysis techniques for complex radiation environments is an ongoing focus in nuclear instrumentation and monitoring. This proposal presents research and development of a generalized real-time neutron monitoring and analysis system, applicable to any detector capable of producing time-tagged neutron count data. While the work is demonstrated using the Neutron Multiplication Analysis Detector (NoMAD), a modular 15-tube helium-3 (He-3) array, due to its availability, spatial resolution, and flexible deployment, the methods developed are extensible to other systems, including organic scintillators and fast digital detectors. This research investigates two complementary analytical techniques for real-time characterization of neutron emitting sources: neutron multiplicity estimation based on the Hage-Cifarelli formalism and spatial localization using supervised machine learning applied to spatial count rate patterns. These methods are designed to operate under dynamic, evolving conditions such as fuel debris retrieval or reactor startup, where neutron-emitting material geometries may be partially unknown or changing over time. By integrating statistical neutron emission data with spatial localization, this research aims to develop and evaluate methods for real time neutron monitoring, source characterization, and material verification. Key contributions include implementation of a low-latency data pipeline for continuous neutron multiplicity analysis, development and validation of machine learning models for spatial inference, and experimental evaluation of system performance under variable measurement conditions. The outcomes are intended to support applications in nuclear safeguards, verification, emergency response, and reactor startup.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

An In Situ , Automated High-Explosives Aging Method Utilizing Two-Dimensional Gas Chromatography–Mass Spectrometry

Understanding chemical changes that occur in high explosives as they age is of great importance to the safe employment and storage of these compounds. Traditional methods of aging high explosives even under accelerated aging conditions are time intensive with durations on the order of months to years. The nature of traditional aging analyses reduces each sample to a snapshot data point often separated widely in time, requiring many assumptions as to how the degradation products develop. Further complicating matters, several analytical techniques are typically employed for each sample analysis in order to ascertain an entire picture of the decomposition pathways. To address these shortcomings with existing methods, a new method of accelerated aging of high explosives utilizing comprehensive two-dimensional gas chromatography coupled to high-resolution mass spectrometry (GC × GC-HRMS) was developed using 2,4,6,8,10,12-hexanitro-2,4,6,8,10,12-hexaazaisowurtzitane (CL-20) as a model compound for method development. This in situ automated method reduces the time scale of aging to a matter of hours using the inlet of the GC × GC as the aging vessel. GC × GC in combination with HRMS allowed for the collection of both evolved gases and other decomposition products produced during the entire aging process in real time with HRMS providing far greater certainty in identification of explosives aging products. Additionally, this method allowed for a higher throughput of samples with greatly simplified sample preparation. Chemometric analysis of the GC × GC-HRMS data set via the alteration analysis (ALA) enabled discovery of statistically significant chemical changes providing insight into the variation of decomposition pathways with varying aging temperatures.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Synthetic spectra for Lyman- α forest analysis in the Dark Energy Spectroscopic Instrument

Synthetic data sets are used in cosmology to test analysis procedures, to verify that systematic errors are well understood and to demonstrate that measurements are unbiased. In this work we describe the methods used to generate synthetic datasets of Lyman-α quasar spectra aimed for studies with the Dark Energy Spectroscopic Instrument (DESI). In particular, we focus on demonstrating that our simulations reproduces important features of real samples, making them suitable to test the analysis methods to be used in DESI and to place limits on systematic effects on measurements of Baryon Acoustic Oscillations (BAO). We present a set of mocks that reproduce the statistical properties of the DESI early data set with good agreement. Additionally, we use a synthetic dataset to forecast the BAO scale constraining power of the completed DESI survey through the Lyman-α forest.

79 ASTRONOMY AND ASTROPHYSICS↗

Topological Signatures of Adversaries in Multimodal Alignments

Topological Data Analysis for Adversarial Detection (LANL O4937) - Detects adversarial examples in vision-language models using persistent homology and two-sample testing. Combines TDA features from CLIP embeddings with statistical methods (ME, SCF, SAMMD, C2ST) for robust detection across ImageNet, CIFAR-10/100.

Bhattarai, Manish↗