Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “analysis workflow”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

ATLAS Data Analysis using a Parallel Workflow on Distributed Cloud-based Services with GPUs

A new type of parallel workflow is developed for the ATLAS experiment at the Large Hadron Collider, that makes use of distributed computing combined with a cloud-based infrastructure. This has been developed for a specific type of analysis using ATLAS data, one popularly referred to as Simulation-Based Inference (SBI). The JAX library is used for the parts of the workflow to compute gradients as well as accelerate program execution using just-in-time compilation, which becomes essential in a full SBI analysis and can also offer significant speed-ups in more traditional types of analysis.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Enabling discovery data science through cross-facility workflows

Experimental and observational instruments for scientific research (such as light sources, genome sequencers, accelerators, telescopes and electron microscopes) increasingly require High Performance Computing (HPC) scale capabilities for data analysis and workflow processing. Next-generation instruments are being deployed with higher resolutions and faster data capture rates, creating a big data crunch that cannot be handled by modest institutional computing resources. Often these big data analysis pipelines also require near real-time computing and have higher resilience requirements than the simulation and modeling workloads more traditionally seen at HPC centers. While some facilities have enabled workflows to run at a single HPC facility, there is a growing need to integrate capabilities across HPC facilities to enable cross-facility workflows, either to provide resilience to an experiment, increase analysis throughput capabilities, or to better match a workflow to a particular architecture. In this paper we describe the barriers to executing complex data analysis workflows across HPC facilities and propose an architectural design pattern for enabling scientific discovery using cross-facility workflows that includes orchestration services, application programming interfaces (APIs), data access and co-scheduling.

Antypas, Katerina B.↗

Enhancement of PyARC for Westinghouse Electric Company’s Lead Fast Reactor Design and Modeling (Final TCF Report)

Westinghouse Electric Company is a nuclear reactor vendor headquartered in the U.S. that is developing advanced reactor technology for the U.S. and global markets. Westinghouse has been relying on the neutronics Argonne Reactor Codes (ARC) executed through the NEAMS Workbench and its PyARC module that are developed under the DOE-NE Nuclear Energy Advanced Modeling and Simulation (NEAMS) and Advanced Reactor Technology (ART) – Fast Reactor programs. Through this user experience, Westinghouse identified several enhancements that would benefit the ARC codes’ usability by the US industry and therefore its commercialization potential. The enhancements were proposed to deliver both improvements in workflow and analysis capabilities to better support effective fast reactor core design and analysis to the nuclear industry. The PyARC workflow was extended in this project by integrating non-neutronic ARC codes DASSH and NUBOW-3D. The Ducted Assembly Steady-State Heat equation (DASSH) code is developed at ANL to perform steady-state thermal hydraulic sub-channel analysis in liquid metal fast reactor assemblies to determine optimized coolant flow and temperature distributions, which in this project was updated and validated for lead fast reactor (LFR) applications. The interface between REBUS and NUBOW-3D were improved in this project to assess the impact of the core restraint design and thermal induced expansion effects on the reactivity of the core, and to model the deformations of the fuel assemblies induced by temperature and irradiation. Finally, the ARC models that were extensively verified and validated through various SFR-based modeling benchmarks are extended in this project through code-to-code comparison on relevant LFR-specific neutronics benchmarks against Monte-Carlo neutronic solutions. Overall, this work enables verification of the capability of the ARC codes for a wide range of Generation-IV reactor designs. The outcome of this project is the release of a comprehensive modeling toolkit of validated, robust and efficient codes, as well as their user interface, that enables industry to perform a wide range of fast reactor analyses for design and licensing of their concepts.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

A Parallel Machine Learning Workflow for Neutron Scattering Data Analysis

As part of a larger effort, this work-in-progress reports the possible advantages of modifying conventional workflows used to generate labelled training samples and train machine learning (ML) models on them. We compare results from three different workflows using neutron scattering data analysis as the motivating application and report about 20% improvement in speedup, with no appreciable loss of model accuracy, over a baseline workflow.

Wang, Tianle↗

Automated Immunoprecipitation Workflow for Comprehensive Acetylome Analysis

Immunoprecipitation is one of the most effective methods for enrichment of lysine-acetylated peptides for comprehensive acetylome analysis using mass spectrometry. Manual acetyl peptide enrichment method using non-conjugated antibodies and agarose beads has been developed and applied in various studies. However, it is time consuming, and can introduce contaminants and variability that leads to potential sample loss and decreased sensitivity and robustness of the analysis. Here we describe a fast, automated enrichment protocol that enables reproducible and comprehensive acetylome analysis using a magnetic bead-based immunoprecipitation reagent.

Lysine acetylation, Acetylome, Acetyl peptide enri↗

HiFiAdapterFilt, a memory efficient read processing pipeline, prevents occurrence of adapter sequence in PacBio HiFi reads and their negative impacts on genome assembly

Abstract Background Pacific Biosciences HiFi read technology is currently the industry standard for high accuracy long-read sequencing that has been widely adopted by large sequencing and assembly initiatives for generation of de novo assemblies in non-model organisms. Though adapter contamination filtering is routine in traditional short-read analysis pipelines, it has not been widely adopted for HiFi workflows. Results Analysis of 55 publicly available HiFi datasets revealed that a read-sanitation step to remove sequence artifacts derived from PacBio library preparation from read pools is necessary as adapter sequences can be erroneously integrated into assemblies. Conclusions Here we describe the nature of adapter contaminated reads, their consequences in assembly, and present HiFiAdapterFilt, a simple and memory efficient solution for removing adapter contaminated reads prior to assembly.

59 BASIC BIOLOGICAL SCIENCES↗

Near real-time streaming analysis of big fusion data

Experiments on fusion plasmas produce high-dimensional data time series with ever-increasing magnitude and velocity, but turn-around times for analysis of this data have not kept up. For example, many data analysis tasks are often performed in a manual, ad-hoc manner some time after an experiment. In this article, we introduce the Delta framework that facilitates near real-time streaming analysis of big and fast fusion data. By streaming measurement data from fusion experiments to a high-performance compute center, Delta allows computationally expensive data analysis tasks to be performed in between plasma pulses. This article describes the modular and expandable software architecture of Delta and presents performance benchmarks of individual components as well as of an example workflow. Focusing on a streaming analysis workflow where electron cyclotron emission imaging (ECEi) data is measured at KSTAR on the National Energy Research Scientific Computing Center's (NERSC's) supercomputer we routinely observe data transfer rates of about 4 Gigabit per second. In NERSC, a demanding turbulence analysis workflow effectively utilizes multiple nodes and graphical processing units and executes them in under 5 min. We further discuss how Delta uses modern database systems and container orchestration services to provide web-based real-time data visualization. For the case of ECEi data we demonstrate how data visualizations can be augmented with outputs from machine learning models. Here, by providing session leaders and physics operators, results of higher-order data analysis using live visualizations may make more informed decisions on how to configure the machine for the next shot.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Apparatus and method for safety analysis evaluation with data-driven workflow

An apparatus and method for system safety analysis evaluation is provided, the apparatus including processing circuitry configured for generating a calculation matrix for a system, generating a plurality of models based on the calculation matrix, performing a benchmarking or convolution analysis of the plurality of models, identifying a design envelope based on the benchmarking or convolution analysis, deriving uncertainty models from the benchmarking or convolution analysis, deriving an assessment judgment based on the uncertainty models and acceptance criteria, defining one or more limiting scenarios based on the design envelope, and determining a safety margin in at least one figure-of-merit for the system based on the design envelope and the acceptance criteria.

Martin, Robert P.↗

Adaptive elasticity policies for staging-based in situ visualization

In situ processing aims to alleviate the growing gap between computation and I/O capabilities by performing data processing close to the data source. In situ processing is widely used to process data generated by multiple data sources, including observation data from edge devices or scientific observational facilities and the simulation data generated by scientific computation on a high-performance computing (HPC) platform. For a scientific workflow that is run on an HPC platform and composed of a simulation program and an in situ data analytics or visualization (abbreviated as ana/vis) task, there is an implicit assumption that the computing resources assigned to the workflow keep static during the workflow execution. However, with the converging trend between the HPC and cloud computing platform, running the in situ ana/vis task in an elastic way is promising to decrease its overhead and improve its resource utilization rate. Resource elasticity represents the ability to change resource configurations such as the number of computing nodes/processes during workflow execution. An elastic job may dynamically adjust resource configurations; it may use a few resources at the beginning and more resources toward the end of the job when interesting data appear. However, it is hard to predict a priori how many computing nodes/processes need to be added/removed during the workflow execution to adapt to changing workflow needs. How to efficiently guide elasticity operations, such as growing or shrinking the number of processes used for in situ analysis during workflow execution, is an open-ended research question. In this article, we present adaptive elasticity policies that adopt workflow runtime information collected during workflow execution to predict how to trigger the addition/removal of processes in order to minimize in situ processing overhead. Taking in situ visualization tasks as an example, we integrate the presented elasticity policies into a staging-based elastic workflow and evaluate its efficiency in multiple elasticity scenarios. Compared with the situation without elasticity or with a static elasticity policy that uses a fixed number of processes for each rescaling operation, the adaptive elasticity policy can save overhead in finding a proper resource configuration and improve resource utilization efficiency. Furthermore, one experiment illustrates that the adaptive elasticity policy saves 41% of core-hours compared with the situation without the resource elasticity.

97 MATHEMATICS AND COMPUTING↗

Assessing the numerical stability of physics models to equilibrium variation through database comparisons on DIII-D

High fidelity kinetic equilibria are crucial for tokamak modeling and analysis. Manual workflows for constructing kinetic equilibria are time consuming and subject to user error, motivating development of automated equilibrium reconstruction tools to provide accurate and consistent reconstructions for downstream physics analysis. These automated tools also provide access to kinetic equilibria at large database scales, which enables the quantification of general uncertainties arising from equilibrium reconstruction techniques. In this paper, we compare a large database of DIII-D kinetic equilibria generated manually by physics experts to equilibria from automated kinetic reconstruction tools, assessing the impact of reconstruction method on equilibrium parameters and resulting magnetohydrodynamic stability calculations. We find agreement among scalar parameters, whereas profile quantities, such as the bootstrap current, show larger disagreements. We analyze ideal kink and classical tearing stability with DCON and STRIDE respectively, finding that the kink stability calculation is generally more robust than the tearing index Δ' calculation. We find that in 90% of cases, both kink stability classifications are unchanged between the manual expert and automated kinetic equilibria.

CAKE↗

Assessing Low-Temperature Geothermal Play Types: Relevant Data and Play Fairway Analysis Methods

The U.S. Department of Energy (DOE) Geothermal Technologies Office (GTO) is supporting the Geothermal Heating and Cooling Geospatial Datasets and Analysis project conducted by the National Renewable Energy Laboratory (NREL) as part of a broader effort to demonstrate the multi-faceted value of integrating geothermal power and geothermal heating and cooling (GHC) technologies into national decarbonization plans and community energy plans. Currently, there is a need to establish baseline low-temperature geothermal resource datasets and evaluate methods for deploying these technologies to provide the basis for supporting private sector investment. This project is focused on collecting baseline datasets, updating conceptual models, and creating Play Fairway Analysis (PFA) workflows for low-temperature (<150 degrees Celsius) geothermal resources of different geothermal play types (i.e., sedimentary basin, orogenic belts, and radiogenic geothermal play types) that could be used for geothermal heating and cooling (GHC), combined heat and power (CHP), and other geothermal direct uses (GDU) applications. Low-temperature geothermal resources are defined as reservoirs - natural or engineered - with temperatures <150 degrees Celsius. While the focus in the NREL effort is on GHC, resources at the upper end of this temperature range can also be used for small-scale power generation. This project does not include Ground Source Heat Pumps (GSHPs) technologies because they can be effectively developed almost anywhere. Low-temperature geothermal resources have not been studied as extensively as higher- to medium-temperature geothermal resources, but there is recent interest in improving understanding of these types of resources with an uptick of interest in geothermal technologies for decarbonizing heating and cooling systems. In addition, Enhanced Geothermal Systems (EGS) and other emerging technologies for exploiting petrothermal resources have opened the possibility of utilizing deep sedimentary basin systems, where porous media provide permeability and high temperatures can be reached at great depths. This project takes the approach of classifying low- temperature geothermal resources by geothermal play type (GPT). We defined and characterized three major classes of low-temperature GPT: sedimentary basins, orogenic systems, and radiogenic systems. We develop methodologies for evaluating and analyzing the potential for these resources building off the PFA approach to de-risking geothermal exploration and characterization. The proposed PFA approach for low-temperature geothermal resources includes: 1) identifying relevant data (e.g., datasets such bottom-hole temperatures from oil and gas wells, heat flow data, Quaternary faults and stress field data, geophysical data, etc.); 2) grouping and weighting of relevant datasets into PFA criteria (e.g., geological, risk, and economic criteria); 3) uncertainty quantification; 4) developing favorability or common risk maps for low-temperature geothermal resources to identify potential locations for more focused data collection; and 5) estimating electric power generation and heating potential at those locations using the GeoRePORT Resource Size Assessment Tool (RSAT). This project should facilitate future deployment of GHC, CHP, and GDU by providing data, tools, and a workflow applicable to low-temperature geothermal resources. Increased deployment of GHC and GDU will help achieve national and local decarbonization goals.

15 GEOTHERMAL ENERGY↗

Frontier Job-Centric Telemetry Dataset

Comprehensive analysis of high-performance computing (HPC) systems requires linking workload execution to system behavior. This kind of analysis is vital for diagnosing performance issues, managing capacity, detecting anomalous workloads, and understanding how applications interact with system hardware. This job-centric telemetry dataset unifies scheduler job records with node-level measurements, enabling direct association between workloads and their corresponding power, thermal, and performance characteristics. It contains sanitized, scheduler related metadata for 152,400 individual jobs that ran on the Frontier supercomputer and ended on selected days throughout 2024 and 2025, a subpopulation of ~6.8% of the total number of allocated jobs with non-zero run time on the system over that same period. Each is linked with files that contain telemetry time series records of the power utilization and temperature behavior of its allocated nodes and their processors during the run time of the job. Where available, a portion of the job files also contain network performance time series. Jobs are sampled from select days that reflect normal levels of user activity and possess job size distributions with large numbers of leadership class jobs (>20% of Frontier nodes). Jobs in this dataset attempt to best represent successful user workflows.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Multiscale Characterization of Additive Manufacturing Components with Computed Tomography, 3D X-ray Microscopy, and Deep Learning

Additive manufacturing (AM) facilitates the creation of complex-geometry parts, driving advancements in lightweight aerospace components, high-efficiency engine cooling channels, and customized medical implants. However, ensuring the quality and reliability of AM parts remains challenging due to internal defects, surface irregularities, porosity, and residual trapped powder, which are often inaccessible to traditional inspection methods. Recent developments in X-ray computed tomography (XCT) and 3D X-ray microscopy (XRM), particularly systems equipped with resolution-at-a-distance (RaaD™) capabilities, enable high-resolution, non-destructive evaluation of AM components across multiple scales, from sub-micrometer to macroscopic levels. This paper explores modern XCT and XRM techniques for multiscale characterization of AM parts, focusing on their ability to detect and analyze defects such as porosity, cracks, inclusions, and surface roughness, while offering insights into defect formation mechanisms, material properties, and process-induced variations. The integration of deep learning (DL) frameworks, including Simurgh, DeepRecon, and DeepScout, enhances XCT/XRM workflows by reducing scan times, improving resolution recovery, and enabling accurate defect detection even with limited projection data. These DL-based methods overcome limitations of traditional reconstruction techniques, enabling faster, more reliable characterization of dense materials like Inconel 718 and novel alloys such as AlCe. Applications include process parameter optimization, high-throughput quality control, and multistage AM process evaluation, with DL-enhanced workflows accelerating analysis times from weeks to days. Correlative imaging approaches further validate XCT and XRM data against scanning electron microscopy (SEM) images of physically sectioned samples, confirming the accuracy of DL-based reconstructions and enabling comprehensive defect analysis. While challenges remain in generalizing DL models to diverse materials and imaging conditions, improvements in resolution, noise reduction, and defect detection highlight the transformative potential of these methods. This multiscale and correlative approach enables precise identification and correlation of microstructural features with the overall performance of AM components. By integrating advanced XCT, XRM, and DL techniques, this paper demonstrates a significant leap forward in AM characterization, offering valuable insights into the relationships between processing parameters, microstructure, and part performance, and driving innovations that enhance the quality and reliability of AM products for demanding industrial applications.

Additive manufacturing↗

Preparation of the Multi-Site Data Processing at the Vera C. Rubin Observatory

The Vera C. Rubin Observatory’s Legacy Survey of Space and Time (LSST) Camera is scheduled to start taking data in the summer of 2025. The Data Release Production will run the LSST Science Pipe software at data facilities in the US, France and the UK. The LSST Science Pipeline consists of complex directed acyclic graphs (DAGs) of tasks. Rubin will use the Production and Distributed Analysis (PanDA) workflow and workload management system to orchestrate this complex workflow and the distribution of workloads to the data facilities. When run end-to-end by a team of data production staff, this processing (the Science Pipelines, distributed by the workflow and workload management system) is referred to as a 'campaign'. This paper describes the central services and data facility specific services that support this multi-site data process model, including the service deployment infrastructure, the workload and workflow system, the Campaign Management tools, and connection to Rubin Data Management. This paper will also mention the experience of processing the Rubin Commissioning Camera data. All these are part of the effort to scale up the processing capabilities for the expected very large data volume from the LSST Camera.

Yang, Wei [SLAC]↗

Linear simulation of magnetohydrodynamic plasma response to three-dimensional magnetic perturbations in high-β P plasmas

In this work, we report the numerical analyses of linear magnetohydrodynamics (MHD) plasma response to applied three-dimensional magnetic perturbations (MPs) in a joint DIII-D/EAST collaboration on high-β P (poloidal beta) plasmas, utilizing the extended-MHD code M3D-C1, with the purpose of realizing a better understanding of the existing experiment in which the n=3 MPs were applied to such high-β P plasmas attempting to control large amplitude type-I ELMs. Such high-β P plasmas obtained at the DIII-D tokamak feature an upper-biased double null configuration, a high edge safety factor q 95 ~6.4, and a stable internal transport barrier (ITB) leading to relatively high core pressures. Single-fluid simulations show that the plasma response to n=3 MPs, including both non-resonant/kinking and resonant components, is significantly weaker than that to n=1 or 2 MPs. To survey the impact of q 95 on plasma response to applied MPs, the SEGWAY (Self-consistent Equilibrium Generating Workflow for AnalYsis) module, developed in the OMFIT integrated modelling framework, is employed to generate a series of equilibria with a wide range of q 95 while other key parameters including the normalized beta, electron density at pedestal top, and plasma shape are kept fixed. Compared to the vacuum response, single-fluid M3D-C1 simulations predict a much more significant decrease of resonant plasma response to the applied n=3 MPs at the maximum penetration radii as q 95 increases. In contrast to single-fluid simulation results showing resonant penetration occurs only near the pedestal top where the E×B toroidal rotation frequency is zero, two-fluid simulations show two comparable resonant penetrations locating near the pedestal top and the ITB foot, where the perpendicular electron rotation frequency is zero. Such resonant field penetration near the ITB foot may be responsible for the observed formation of a staircase structure in both electron density and temperature profiles and thereby a considerable deterioration of global plasma performance when MPs are applied in high-β P plasmas. Motivated by this numerical work, we provide some ideas for the future research, with the purpose of realizing effective ELM control in such high-β P plasmas on the DIII-D and EAST devices.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗