Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “continuous integration”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

NMSBA: Continuous Application Benchmarking & Analysis – CABA

The CABA project is a test of a new tool, Survey,developed by Trenza to be used not only for benchmarking or profiling programs but also to allow incorporation of the information provided by Survey to be utilized in a CI, continuous integration,tool such as GitLab CI.Survey is foremost a means of assessing code performance in terms of time and operations which for computer programmers is known as benchmarking.

97 MATHEMATICS AND COMPUTING↗

Accomplishments of Sandia and Kitware CMake/CTest/CDash Contract for (FY2017-2020)

We describe the accomplishments jointly achieved by Kitware and Sandia over the fiscal years 2016 through 2020 to benefit the Advanced Scientific Computed (ASC) Advanced Technology Development and Mitigation (ATDM) project. As a result of our collaboration, we have improved the Trilinos and ATDM application developer experience by decreasing the time to build, making it easier to identify and resolve build and test defects, and addressing other issues . We have also reduced the turnaround time for continuous integration (CI) results. For example, the combined improvements likely cut the wall clock time to run automated builds of Trilinos posting to CDash by approximately 6x or more in many cases. We primarily achieved these benefits by contributing changes to the Kitware CMake/CTest/CDash suite of open source software development support tools. As a result, ASC developers can now spend more time improving code and less time chasing bugs. And, without this work, one can argue that the stabilization of Trilinos for the ATDM platforms would not have been feasible which would have had a large negative impact on an important internal FY20 L1 milestone.

97 MATHEMATICS AND COMPUTING↗

FY20 Proxy App Suite Release: Report for ECP Proxy App Project Milestone ADCD-504-10

Version 4.0 of the ECP Proxy App Suite is practically unchanged from the previous release. The current set of proxies has proven useful for many aspects of benchmarking and co-design and we see little reason to alter the suite. Although there have been few changes to the ECP suite, the team has been hard at work in other areas. In the area of Machine Learning (ML) we have now created a separate proxy suite dedicated to this scientific applications of ML. The suite includes: miniGAN (Generative Adversarial Networks), miniRL (Reinforcement Learning), CRADL (inline inference), Cosmoflow-Benchmark (Convolutional Neural Network), and MLPerf-DeepCam (Climate Segmentation Benchmark). Section 3 contains more information about these proxies as well as the principles that are guiding the development of the suite. We have surveyed available proxies for several application domains including Computational Fluid Dynamics, Quantum Chemistry, Quantum Computing Simulation, Molecular Dynamics, Monte Carlo Transport, and Density Functional Theory to identify gaps in proxy coverage. Several new proxy apps are either already available or will be released soon to fill these gaps. Section 4 provides full details. Finally, section 5 reports on our continued collaboration with the ECP Continuous Integration (CI) effort to use proxy apps to help identify problems and roadblocks to cross-lab CI. We have also assisted the El Capitan Center of Excellence (COE) to stand up a CI system that will be used to test releases of HPE and AMD software stacks. We hope that the COE effort can serve as a model for ECP by showing how the Proxy App Team can work with AD and ST teams identify critical features, kernels, patterns, etc. and incorporate them into a CI system that will help ensure that Frontier and Aurora will provide those needed capabilities.

97 MATHEMATICS AND COMPUTING↗

High Temperature Steam Electrolysis Process Performance and Cost Estimates

Technology readiness levels (TRLs) of electrolysis systems have dramatically increased in recent years as the interest in clean hydrogen production and decarbonization of transportation, industrial and other sectors increases across the globe. This is especially true of high temperature steam electrolysis (HTSE) / solid oxide electrolysis cell (SOEC) systems which show promise of much higher system efficiencies than other more developed electrolysis technologies. This possibility of higher efficiencies of HTSE / SOEC systems has been previously assumed to be theoretically possible but in recent years it has become less theoretical and more realistic as an increasing amount of suppliers complete lab and pilot tests showing very promising results. Research in the areas of manufacturing techniques, material selection, electrode and electrolyte compositions, and balance of plant size and integration continues at a fast pace as an increasing number of suppliers both internationally and domestically become involved. The advantages of HTSE become more pronounced when HTSE is coupled with nuclear power plants (NPPs). This is because thermal energy produced by the nuclear reactor can be used in a series of heat transfer loops and heat exchangers to vaporize HTSE feedwater, which drastically improves the economics of the process. Idaho National Laboratory (INL) has been very involved in the research and modeling of HTSE systems for a number of years, in collaboration with other national laboratories, academia, and industry stakeholders both on the hydrogen production as well as the hydrogen demand side. The modeling completed over the years on a large variety of projects has led to a wealth of knowledge at INL including in the area of the technoeconomic assessment (TEA) of HTSE systems. TEAs include process modeling of the HTSE systems to calculate system energy requirements and equipment sizing, followed by estimation of capital and operating costs to enable calculation of the levelized cost of hydrogen (LCOH). The TEA work performed has produced incremental improvements and tuning of the methods, assumptions, models, and results of the analyses as well as providing some opportunities for validating these results. The purpose of this document is to record the current baseline HTSE analyses led by INL to show the current status of assumptions and costs of these systems. Given the rapid development of this technology, the variety of suppliers entering the space, and the increasing attention government and industry are giving to such systems, this document may be updated on a periodic basis with updated analysis and assumptions. This document compiles various analyses results and approaches completed over a period of years into a single document to be used as a baseline going forward. It represents what the INL HTSE analysis group assumes to be the internal best estimate of the current operation, costs, and landscape of the HTSE industry state of the art capability for current SOEC technology in an Nth-of-a-Kind (NOAK) plant, which in this study is defined as existence of the manufacturing capacity to support previous deployment of N = 100 count of 25 MWe modular HTSE blocks (with modular equipment component cost reductions specified as following a 95% learning curve). That said, This is a public document and as such so no proprietary data was used or included in this report. There may be HTSE suppliers that have performance specifications, and cost estimates, and test data that differ from the analysis presented in this document. This document is meant to be a best conservative estimate of the technology and not an absolute reference.

08 HYDROGEN↗

Expanding the representation of aerosol, cloud, and precipitation processes with graph network-based simulators

We explored a novel framework for simulating the small-scale processes that drive the evolution of aerosol, cloud, and precipitation particles, which are a critical gap in the predictive understanding of weather and climate. Particle-based methods have emerged as an effective tool for modeling aerosol-cloud-precipitation interactions, but existing particle-based models are computationally too expensive to simulate the large domains relevant for the atmosphere or to represent the full suite of relevant processes. The lack of a comprehensive and efficient reference model is a critical bottleneck in our understanding of cloud and precipitation processes and our ability to parameterize these processes for regional- and global-scale simulations. To address this need, we explored an approach to accelerate and expand particle-based models using a new machine learning approach, graph network-based simulators (GNS). Rather than modeling the evolution of the system by numerically integrating continuity equations, the GNS represents dynamics through learned message passing. Our aim was to develop fast and accurate surrogate models for particle-based simulations. We explored applying GNS to simulate cloud droplet transport, growth, and evaporation under turbulent conditions, but we found the GNS over-smoothed the simulations. We then applied the GNS to simulate aerosol dynamics through gas condensation and found the GNS was able to reproduce the benchmark, physics-based simulation with high accuracy.

54 ENVIRONMENTAL SCIENCES↗

Fiscal Year 2024 Software Quality Assurance Activities for the ARC Software

The continued goal of the ARC SQA project in the Advanced Reactor Technologies program of DOE is to resolve the QA gaps for the ARC software that limit, or prevent, commercialization of the software for industry users. This project started in earnest in fiscal year 2023 which saw the entire code system moved from a SVN repository to a GitLab repository and an associated software quality assurance plan (SQAP) developed and ratified. Most of the QA gaps in the ARC software were identified in collaboration with industry partners and work begin in fiscal year 2023 and continued in 2024. The primary documentation that is missing includes user manuals, user guides, software verification reports, and code coverage assessments. The SUMMAR manual was completed this fiscal year and work was started on creating manuals for SE2ANL, SE2RCT, and DASSH. Software verification work was carried out for DIF3D and REBUS in a previous program and the current fiscal year saw the completion of software verification reports for GAMSOR, GAMSRC, VARPOW, EvaluateFlux, and SUMMAR. The goal for the next fiscal year is to complete the PERSENT software verification work and begin planning the software verification work for DASSH, SE2ANL, and SE2RCT. The code coverage reports for DIF3D and MC2-3 were completed in the previous fiscal year and the goal is to generate code coverage reports for REBUS, GAMSOR, PERSENT, and DASSH in the coming fiscal year. A considerable amount of effort was spent in the current fiscal year working on the continuous integration capability for automated regression testing in GitLab. The first version of the testing was created in the previous fiscal year and applied to DIF3D and its utility programs. That testing was extended this year to cover GAMSOR, REBUS, and PERSENT. To accomplish this, the first version of the new testing methodology had to be updated to make a single output checking methodology viable for all of the ARC software. This will result in a single document to detail the automated regression testing methodology and minor documents to detail the tolerance settings that have been applied to the output for each ARC code. The previous methodology put into place with SVN would have required a separate document for each ARC code to detail the output checking methodology and the tolerance settings for the output from each code. Because some of our industry partners are providing funds to add new capabilities to the ARC software to meet their needs, all of which must be reviewed and approved by the SQA program funded by this project, a summary of that development work is detailed in this report. Overall progress on resolving the QA gaps has been good this year with the most impactful improvement for our industry partners in capability being the creation of a threaded version of DIF3D-VARIANT that allows the DIF3D, REBUS, and GAMSOR run times to be reduced by a factor of 4-6. The most impactful QA gap that was resolved was the software verification of GAMSRC and VARPOW.

97 MATHEMATICS AND COMPUTING↗

On the use of Graphs for Test Sequence Selection

This report demonstrates that applying graph theory techniques provides a way to obtain sufficient statistics in finding errors when testing complex state machines. It discusses how to define the tests, then demonstrates how to automatically generate test suites that diversify test cases, subject to constraints. If included within a continuous integration approach, these constructs provide an unbiased means to systematically check for errors within the latest controller software release.

97 MATHEMATICS AND COMPUTING↗

High-Burnup LOCA Burst Susceptibility BISON Analysis in PWRs and BWRs

Accurately assessing high-burnup fuel behavior during loss-of-coolant accidents (LOCAs) is essential for understanding fuel fragmentation, relocation, and dispersal (FFRD) risks across the US light-water reactor fleet. This work updates previous Nuclear Energy Advanced Modeling and Simulation (NEAMS) Program multiphysics LOCA analyses for a pressurized water reactor (PWR) and a boiling water reactor (BWR) by incorporating recent model and material property advancements in the BISON fuel performance code, including a high-burnup structure (HBS) model, revised cladding burst criteria, and updated thermal–mechanical correlations. This update was needed to support ongoing industry initiatives and upcoming regulatory changes. Full-core, rod-resolved operating histories generated using Virtual Environment for Reactor Analysis (VERA) and system-level LOCA conditions obtained from TRACE were applied to statistically representative rod samples in BISON to evaluate burst behavior and FFRD susceptibility. These calculations used two cladding burst correlations and three fuel pulverization models so that the predictions of these models could be compared. The updated PWR simulations show markedly improved numerical stability as the number of crashed simulations decreased by 95% compared to the previous study, and hence higher confidence in results. The updated PWR simulations predicted cladding bursts exclusively among once-burned, high-power rods, with two different cladding burst models identifying the same burst-susceptible population. Resulting FFRD susceptibility estimates are significantly reduced compared with earlier studies, driven by cooler predicted fuel and plenum temperatures, lower hoop strains, and reduced fission gas release in the updated models. In contrast, none of the BWR rods were predicted to burst under either burst criterion, reaffirming minimal BWR FFRD susceptibility even with updated HBS and material models. Comparisons between the PWR and BWR end-of-cycle predictions are made. Comparison with prior work highlights significant shifts in PWR fuel performance metrics and confirmation of earlier BWR conclusions. Overall, the updated results underscore the importance of having high-resolution detailed modeling capability and continuously integrating evolving material models and physics into high-resolution multiphysics simulations. The unified assessment presented here strengthens confidence in predicting high-burnup LOCA behavior by improving agreement between different cladding burst correlations. These results also provide an improved foundation for future BISON model development, FFRD susceptibility calculations.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Containerization in ATLAS Software Development and Data Production

The ATLAS experiment's software production and distribution on the grid benefits from a semi-automated infrastructure that provides up-to-date information about software usability and availability through the CVMFS dis-tribution service for all relevant systems. The software development process uses a Continuous Integration pipeline involving testing, validation, packag-ing and installation steps. For opportunistic sites that can not access CVMFS, containerized releases are needed. These standalone containers are currently created manually to support Monte-Carlo data production at such sites. In this paper we will describe an automated procedure for the containerization of AT-LAS software releases in the existing software development infrastructure, its motivation, integration and testing in the distributed computing system.

97 MATHEMATICS AND COMPUTING↗

High Temperature Steam Electrolysis Process Performance and Cost Estimates - DOE Hydrogen Program AMR Presentation

Technology readiness levels (TRLs) of electrolysis systems have dramatically increased in recent years as the interest in clean hydrogen production and decarbonization of transportation, industrial and other sectors increases across the globe. This is especially true of high temperature steam electrolysis (HTSE) / solid oxide electrolysis cell (SOEC) systems which show promise of much higher system efficiencies than other more developed electrolysis technologies. This possibility of higher efficiencies of HTSE / SOEC systems has been previously assumed to be theoretically possible but in recent years it has become less theoretical and more realistic as an increasing amount of suppliers complete lab and pilot tests showing very promising results. Research in the areas of manufacturing techniques, material selection, electrode and electrolyte compositions, and balance of plant size and integration continues at a fast pace as an increasing number of suppliers both internationally and domestically become involved. The advantages of HTSE become more pronounced when HTSE is coupled with nuclear power plants (NPPs). This is because thermal energy produced by the nuclear reactor can be used in a series of heat transfer loops and heat exchangers to vaporize HTSE feedwater, which drastically improves the economics of the process. Idaho National Laboratory (INL) has been very involved in the research and modeling of HTSE systems for a number of years, in collaboration with other national laboratories, academia, and industry stakeholders both on the hydrogen production as well as the hydrogen demand side. The modeling completed over the years on a large variety of projects has led to a wealth of knowledge at INL including in the area of the technoeconomic assessment (TEA) of HTSE systems. TEAs include process modeling of the HTSE systems to calculate system energy requirements and equipment sizing, followed by estimation of capital and operating costs to enable calculation of the levelized cost of hydrogen (LCOH). The TEA work performed has produced incremental improvements and tuning of the methods, assumptions, models, and results of the analyses as well as providing some opportunities for validating these results. The purpose of this document is to record the current baseline HTSE analyses led by INL to show the current status of assumptions and costs of these systems. Given the rapid development of this technology, the variety of suppliers entering the space, and the increasing attention government and industry are giving to such systems, this document may be updated on a periodic basis with updated analysis and assumptions. This document compiles various analyses results and approaches completed over a period of years into a single document to be used as a baseline going forward. It represents what the INL HTSE analysis group assumes to be the internal best estimate of the current operation, costs, and landscape of the HTSE industry state of the art capability for current SOEC technology in an Nth-of-a-Kind (NOAK) plant, which in this study is defined as existence of the manufacturing capacity to support previous deployment of N = 100 count of 25 MWe modular HTSE blocks (with modular equipment component cost reductions specified as following a 95% learning curve). That said, this is a public document and as such so no proprietary data was used or included in this report. There may be HTSE suppliers that have performance specifications, and cost estimates, and test data that differ from the analysis presented in this document. This document is meant to be a best conservative estimate of the technology and not an absolute reference.

08 HYDROGEN↗

Modernization efforts for the R -Matrix code SAMMY [Abstract]

The R-Matrix code SAMMY is a widely used nuclear data evaluation code focused on the resolved range, which includes corrections for experimental effects. The code is still mostly written in Fortran 77, and uses a memory management system suitable for the time of its initial writing (1984). A modernization effort is under way to bring the code in-line with modern software development practices. A continuous-integration testing framework was added, automating the large existing set of test cases. It is run on every commit. The memory management was updated to current standard practices suitable for modern software analysis tools. The code can be obtained from https://code.ornl.gov/RNSD/SAMMY. The resonance parameters and covariance information are now stored in C++ objects shared by SAMMY and AMPX, the processing code that generates nuclear data libraries for SCALE. This allows for easier maintenance and access to the resonance parameters inside and outside of SAMMY. This feature is already used by accessing and changing parameters in memory in the Bayesian Monte Carlo Evaluation Framework for Cross Sections Nuclear Data and Integral Benchmark Experiments project, Further plans include the switch to the ENDF reading and writing routines in AMPX, as these routines are more robust, easier to maintain, and support more features. Of note here is support for the new GNDS format. Previously it wasn’t easy to share the full covariance matrix for evaluations containing more than one isotope due to limitations on the ENDF format; this is now supported in GNDS. The data are currently available in a binary SAMMY format and can be exported to GNDS to make them more widely available and sharable. The next step will be to use the same resonance processing code at 0K in AMPX and SAMMY as one of the available Reich-Moore R-Matrix formalism. The first step toward this goal is to isolate the reconstruction into a module that takes resonance parameters as its input and does not depend on SAMMY global parameters. This goal has been achieved and it should now be possible to more easily change the resonance formalism and add enhancements as the Phenomenological R-Matrix parameterization of direct, doorway, and compound nuclear reactions discussed elsewhere on this conference. This concerted modernization and enhancement effort provides multiple advantages to the nuclear data community. It will allow parameter optimization using enhanced formalisms, including experimental effects, that better match complex experimental data. Then those evaluated parameters can immediately be passed off to AMPX to be reconstructed with the exact same cross section model and be put into a data library for subsequent testing using SCALE and the Valid Benchmark suite or other suitable benchmark suites.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

New Developments and Capabilities Within WEC-Sim: Preprint

WEC-Sim is an open-source software for simulating wave energy converters, which has been actively developed and applied since its initial release in 2014 to simulate a wide variety of device archetypes. WEC-Sim is developed jointly by the National Renewable Energy Laboratory (NREL) and Sandia National Laboratories (Sandia) within the MATLAB/SIMULINK environment. A general wave-to-wire model begins with a deployment site resource characterization, which is used to complete the hydrodynamic simulation of wave energy converters (WEC), with the power generation profile imported to a grid simulator to understand the influence on the local electrical network. While modeling the entire wave-to-wire is difficult and encompasses multiple time scales and physics, WEC-Sim is focused on the hydrodynamics simulation to predict, analyze, and optimize WEC dynamics and power performance. WEC-Sim simulations are performed in the time domain based on the radiation and diffraction method using hydrodynamics coefficients derived from boundary element method (BEM)-based frequency-domain potential flow solvers (e.g., WAMIT, NEMOH, Capytaine, or ANSYS-AQWA). With this level of modeling fidelity, WEC-Sim can handle floating body hydrodynamics, mechanical and electrical power generation methods, advanced control implementation, mooring systems, and other unique applications such as desalination. Additional WEC-Sim functionalities include pre-built Simulink blocks and MATLAB scripts that can simulate a wide range of floating systems and the corresponding auxiliary subsystems. The developers of WEC-Sim continue to release new versions of the software, at least annually, with our latest release in September 2022. These releases include bug fixes, updates to software documentation, as well as new features to expand WEC-Sim's capabilities to model a wide range of WEC concepts. This publication will highlight the new features added to WEC-Sim between versions 4.1.0 to 5.0.1 which spans over a two year period from June 2020 to September 2022. New features to be described will include topics such as continuous integration checks, revised Morison Element and nonlinear hydro implementations, run directly from Simulink (required for hardware-in-the-loop execution), BEMIO updates to import Capytaine BEM hydrodynamics, addition of cable blocks, and new wave visualization features.

TIDAL AND WAVE POWER↗

Controls Status of Fermilab's PIP-II Project

The Fermilab Proton Improvement Project II (PIP-II) is building a new Super Conducting Linear Accelerator (SCL) accelerating protons to 800 MeV for injection into the rest of the FNAL beam complex. Key progress since the last status report given at ICALEPCS includes the adoption of modern DevOps practices with continuous integration and GitOps-based deployments, commissioning of EPICS-based systems at the Cryomodule Test Facility, and integration of a Virtual Accelerator framework for application development ahead of installation. In parallel, web-based applications using Dart and Flutter have matured, providing secure, unified access to both EPICS and legacy ACNET data. Data acquisition and timing systems have also evolved. This paper presents the current state of controls, emphasizing these recent developments and outlining upcoming milestones as PIP-II approaches commissioning of its cryoplant in 2026 and the Warm Front End in 2027.

Crisp, D. B. [Fermilab]↗

Integrated combustor nozzles with continuously curved liner segments

An integrated combustor nozzle includes an inner liner segment; an outer liner segment; and a panel extending radially between the inner and outer liner segments. The panel includes a forward end, an aft end, and a side walls extending axially from the forward end to the aft end. The aft end defines a turbine nozzle having a trailing edge circumferentially offset from the forward end. The inner liner segment has a pair of sealing surfaces, each of which defines a first continuous curve in the circumferential direction. The outer liner segment has a pair of sealing surfaces, each of which defines a second continuous curve in the circumferential direction. In some instances, the curves are monotonic in the circumferential direction. A segmented annular combustor including an array of such integrated combustor nozzles is also provided.

Berry, Jonathan Dwight↗

Final Technical Report - Rapid Surface Microanalysis using a Low Temperature Plasma

This project focused on improving our current understanding and scientific knowledge in the area of plasma-surface interactions and plasma assisted material synthesis related to advanced microelectronics and nanotechnology. Current challenges include: controlling the interaction of Low Temperature Plasma (LTP) with a single layer of atoms to manufacture integrated circuits, continued miniaturization of integrated circuits, LTP processing of material surfaces and thin films to enable industrial scale fabrication of advanced microelectronics, synthesis of new materials, nanomaterials, nanotubes, and complex materials, Technology developed in this subtopic is of value to either (i) enable scans of surfaces (~1 sq. cm area) using various microscopies (electron, optical, other) at high resolution (micron or sub-micron resolution) rapidly (hours or days rather than years to complete a high-resolution scan of such a large surface area), or (ii) enable scans of surfaces (~1 sq. cm area) using various microscopies (electron, optical, other) at relatively low resolution rapidly, then apply algorithms to select spots for micron-scale imaging. Sputtering occurs when particles of a solid material are ejected from its surface by energetic particles from a plasma. While the degradation of the solid material and the subsequent deposition of the ejected material onto vulnerable surfaces are the usual subjects of sputtering studies, plasma science has yet to be combined with sputtering to create new diagnostics devices and systems. Small changes in the design of the plasma discharge device make it possible to create broad plasma beams for rapid scanning or small plasma beams to obtain the distribution of ejected elements with micron resolution. In the high-resolution use, the ion flux is extracted from the gas-discharge plasma and focused by a spherical emission surface to micron sizes onto the target specimen, providing very local sputtering and local elemental analysis. We call this “self-focusing”. The radiation from the excited and ionized sputtered atoms is recorded by a spectrometer through a window and fiberglass cable and analyzed with standard software packages used for optical glow discharge spectroscopy. Computer simulations of beam formation were used to verify and optimize the designs to be tested. A prototype was designed, constructed, and used to start experiments of beam formation.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Integrated System and Application Continuous Performance Monitoring and Analysis Capability

Scientific applications run on high-performance computing (HPC) systems are critical for many national security missions within Sandia and the NNSA complex. However, these applications often face performance degradation and even failures that are challenging to diagnose. To provide unprecedented insight into these issues, the HPC Development, HPC Systems, Computational Science, and Plasma Theory & Simulation departments at Sandia crafted and completed their FY21 ASC Level 2 milestone entitled "Integrated System and Application Continuous Performance Monitoring and Analysis Capability." The milestone created a novel integrated HPC system and application monitoring and analysis capability by extending Sandia's Kokkos application portability framework, Lightweight Distributed Metric Service (LDMS) monitoring tool, and scalable storage, analysis, and visualization pipeline. The extensions to Kokkos and LDMS enable collection and storage of application data during run time, as it is generated, with negligible overhead. This data is combined with HPC system data within the extended analysis pipeline to present relevant visualizations of derived system and application metrics that can be viewed at run time or post run. This new capability was evaluated using several week-long, 290-node runs of Sandia's ElectroMagnetic Plasma In Realistic Environments ( EMPIRE ) modeling and design tool and resulted in 1TB of application data and 50TB of system data. EMPIRE developers remarked this capability was incredibly helpful for quickly assessing application health and performance alongside system state. In short, this milestone work built the foundation for expansive HPC system and application data collection, storage, analysis, visualization, and feedback framework that will increase total scientific output of Sandia's HPC users.

97 MATHEMATICS AND COMPUTING↗

Integrated System and Application Continuous Performance Monitoring and Analysis Capability (Final)

Scientific applications run on high-performance computing (HPC) systems are critical for many national security missions within Sandia and the NNSA complex. However, these applications often face performance degradation and even failures that are challenging to diagnose. To provide unprecedented insight into these issues, the HPC Development, HPC Systems, Computational Science, and Plasma Theory & Simulation departments at Sandia crafted and completed their FY21 ASC Level 2 milestone entitled "Integrated System and Application Continuous Performance Monitoring and Analysis Capability." The milestone created a novel integrated HPC system and application monitoring and analysis capability by extending Sandia’s Kokkos application portability framework, Lightweight Distributed Metric Service (LDMS) monitoring tool, and scalable storage, analysis, and visualization pipeline. The extensions to Kokkos and LDMS enable collection and storage of application data during run time, as it is generated, with negligible overhead. This data is combined with HPC system data within the extended analysis pipeline to present relevant visualizations of derived system and application metrics that can be viewed at run time or post run. This new capability was evaluated using several week-long, 290-node runs of Sandia’s ElectroMagnetic Plasma In Realistic Environments (EMPIRE) modeling and design tool and resulted in 1TB of application data and 50TB of system data. EMPIRE developers remarked this capability was incredibly helpful for quickly assessing application health and performance alongside system state. In short, this milestone work built the foundation for expansive HPC system and application data collection, storage, analysis, visualization, and feedback framework that will increase total scientific output of Sandia’s HPC users.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Continued performance improvement and integration of MOOSE's thermal-hydraulics capabilities (M3 Milestone Report)

This work introduces performance, robustness and workflow improvements to Multiphysics Object-Oriented Simulation Environment (MOOSE)-based thermal-hydraulics solvers. It presents work related to the acceleration of segregated fluid dynamics algorithms, which show approximately a factor of 10 speedup compared to the preceding implementation. Additionally, we discuss approaches to use advanced, Schurr complement-based, field split preconditioners for monolithic solution algorithms relying on the finite volume method. The presence of the Rhie-Chow interpolation makes the utilization of this preconditioner challenging, but the results indicate that for a moderately large problem a factor of 3.4 speedup can be achieved in conjunction with a factor of 3.5 reduction in memory usage. Furthermore, we introduce several pseudo-time stepping approaches to MOOSE for the robust convergence to steady-state solutions when steady-state solves don't converge due to the initial guesses being too far from the solution in Newton's method. Every MOOSE-based application has access this algorithm and can benefit from its use. Moreover, several new avenues have been presented for importing meshes from commercial software which make meshing easier. Lastly, the Component system within the Thermal-Hydraulics Module (THM) of MOOSE is abstracted by separating geometry- and physics-related properties.

97 MATHEMATICS AND COMPUTING↗