Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “failure analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 199 records · Page 11

Engineering in Cyber Resilience with Cyber-Informed Engineering

Engineers have super powers to provide cybersecurity resilience with deterministic engineering solutions and to protect systems from the most catastrophic consequences that a cyber saboteur could cause. Come to this session to learn how to use engineering risk management skills to harden your engineered systems from cyberattacks. Objective 1 Identify what system functions could be digitally induced to cause undesired high-impact consequences. Objective 2 Analyze how loss or instability of digital controls in a subsystem could lead to high-impact consequences. Objective 3 Analyze how loss or instability in the digital connectivity between systems could lead to high-impact consequences. Objective 4 Identify engineering controls which could build resilience by eliminating digital loss or instability pathways or reduce the impact of digital loss or instability. This presentation will introduce Cyber-Informed Engineering, described below, and walk participants through specific engineering use cases to show how engineers can consider the potential for cyber sabotage in their existing system designs and enact deterministic engineering-based controls which eliminate pathways for attack or mitigate specific consequences. A wide variety of application use cases will be considered so that audience members can align the material with familiar engineering applications. CIE is an engineering approach that integrates cyber resilience into the conception, design, build, and operation of any physical system that has digital connectivity, sensors, monitoring, or control. CIE offers the opportunity to use engineering to eliminate or mitigate avenues for cyber attack—starting from the earliest stage of design and continuing throughout the system’s lifecycle. Today, engineers and industrial control system (ICS) technicians build engineered systems with specific goals for safety, reliability, and functionality. While systems engineering includes considerable safety and failure mode analysis, cybersecurity risks are often not specifically addressed—particularly the risks of intentional cyber compromise, exploitation, and misuse. Cyber-Informed Engineering pairs well with traditional cyber defenses and offers an extra designed-in protection to eliminate the most catastrophic consequences which can be realized by an adversary should traditional cyber defenses fail.

42 ENGINEERING↗

Hydrogen Component Reliability Database (HyCReD)

The Hydrogen Component Reliability Database (HyCReD) is a collaborative project between the National Renewable Energy Laboratory, the University of Maryland, and hydrogen stakeholders to improve safety reliability for hydrogen facilities by integrating risk reduction methodologies and component reliability data taxonomies that support hydrogen infrastructure failure rate analysis.

availability↗

Addressing Failures in Molten Salt Thermal Energy Storage Tank for Central Receiver Concentrating Solar Power Plants

The thermal energy storage (TES) system is a critical component in concentrated solar power (CSP) plants that increases the plant's capacity factor and economic competitiveness by reducing the levelized cost of energy (LCOE) while simultaneously increasing the value of the delivered energy. Failures including molten-salt leaks and diverse localized cracking after several months to a few years of operation have been reported in hot tanks for CSP plants operating around the world. A model of a molten salt thermal energy storage tank was developed and validated to analyze the impact of different tank design features on the temperature and stress distributions as a function of typical plant operation conditions. Design features included the floor plate thicknesses, friction coefficients between the tank floor and the foundation, and the sparger ring location. Maximum stresses in the tank floor frequently surpassed the yield point of the material during operation, leading to a detriment of the tank's lifetime. Recommendations on design features to improve the reliability of new molten salt tanks for CSP plants are provided.

concentrating solar power↗

Controlled Pyroshock Test Transients Can Be Used To Better Match Operational Shock Transients and Reduce Artificial Pyroshock Test Failures [Slides]

Shock analysis is typically done via SRS (shock response spectrum). Oftentimes shock analysis is not performed because most loads specifications only supply SRS requirements with no representative shock transients that can be used for analysis. SRS analysis is too conservative. There is no structural dynamics involved as all modes are assumed to have peak responses simultaneously and in-phase with each other. It is possible that all other loads and combination of loads show the design to be good, whereas SRS analysis (1000 g’s are involved) results may require redesign. Multiple shock transients can satisfy a SRS, yet most loads documents don’t provide a representative shock transient that could be used for shock analysis. Shock testing is less common than harmonic and random vibration testing. Engineers typically go into shock tests with very little knowledge of what to expect. A pass/fail type of approach with very little engineering done to prepare for the test or to control the type of shock transient is used. Most engineers will not model the shock test fixture due to lack of knowledge about the details of the test setup. Shock analysis and testing is often an open loop type situation with no test setup modeling, no test correlation, and no control over the test shock transient. In this work, if we can create applied test force transient that comes close to matching calculated force transient then a more reasonable acceleration transient will be created at the test reference accelerometer that better matches operational shocks.

42 ENGINEERING↗

Risk Assessment of EIC Central Detector (ePIC) Solenoid Magnet (MARCO)

As part of the BNL-JLab-CEA Electron Ion Collider (EIC) collaboration, the design of a 2 T, 2.8 m bore diameter, 3.8 m long conduction cooled superconducting detector magnet design is completed. Such magnet will be employed at the interaction region of ePIC for physics experiment. The magnet is a passive shielded solenoidal magnet system consisting of 3 coils wound with specially designed conductor using NbTi Rutherford type cable in copper stabilized channel. This paper describes the risk analysis as the part of Failure Modes and Effects Analysis (FMEA) that was carried out as a team to identify their various failure modes and risks associated with the magnets system. In conclusion, this FMEA is intended to become the content of the designed document as an integral part of the engineering assessment and the statement of work for the potential vendors towards design and built.

Ghoshal, Probir K. [Thomas Jefferson National Acce↗

Integrated Transcriptomic and Proteomic Analysis Identifies Plasma Biomarkers of Hepatocellular Failure in Alcohol-Associated Hepatitis

Alcohol-associated hepatitis (AH) is a form of liver failure with high short-term mortality. Recent results have shown that HNF4a defective function and systemic inflammation are major disease drivers of AH. Plasma biomarkers of hepatocyte function could be useful for diagnostic and prognostic purposes. Herein an integrative analysis of hepatic RNAseq and liquid chromatography-tandem mass spectrometry (LC-MS/MS) was performed to identify plasma protein signatures for mild and severe AH patients. Alcohol-related liver disease cirrhosis (ALD)(AC), non-alcoholic fatty liver disease (NALFD), and healthy subjects (HC) were used as comparator groups. Identified proteins primarily involved in hepatocellular function were decreased in AH patients which included hepatokines, clotting factors, complement cascade components, and hepatocyte growth activators. A protein signature of AH disease severity was identified including thrombin (THRB), hepatocyte growth factor alpha (HGFA), clusterin (CLUS), human serum factor H-related protein (FHR1) and kallistatin (KAIN), which exhibited large abundance shifts between severe and non-severe AH. The combination of THRB and HGFA discriminated between severe and non-severe AH with high sensitivity and specificity. These findings were correlated with the liver expression of genes encoding secreted proteins in a similar cohort, finding a highly consistent plasma protein signature reflecting HNF4A and HNF1A functions. This unbiased proteomic-transcriptome analysis identified plasma protein signatures and pathways associated with disease severity, reflecting HNF4A/1A activity useful for diagnostic assessment in AH.

60 APPLIED LIFE SCIENCES↗

CAMERA: A method for cost-aware, adaptive, multifidelity, efficient reliability analysis

Estimating probability of failure in aerospace systems is a critical requirement for flight certification and qualification. Failure probability estimation involves resolving tails of probability distributions, and Monte Carlo sampling methods are intractable when expensive high-fidelity simulations have to be queried. Here, we propose a method to use models of multiple fidelities that trade accuracy for computational efficiency. Specifically, we propose the use of multifidelity Gaussian process models to efficiently fuse models at multiple fidelity, thereby offering a cheap surrogate model that emulates the original model at all fidelities. Furthermore, we propose a novel sequential acquisition function based experiment design framework that can automatically select samples from appropriate fidelity models to make predictions about quantities of interest at the highest fidelity. We use our proposed approach in an importance sampling setting and demonstrate our method on the failure level set and probability estimation on synthetic test functions and two real-world applications, namely, the reliability analysis of a gas turbine engine blade using a finite element method and a transonic aerodynamic wing test case using Reynolds-averaged Navier-Stokes equations. We show that our method predicts the failure boundary and probability more accurately and at a fraction of the computational cost compared with using just a single expensive high-fidelity model. Finally, we show that our sequential approach is guaranteed to asymptotically converge to the true failure boundary with high probability.

97 MATHEMATICS AND COMPUTING↗

Reliability Analysis of Power Grids Considering Component Failures of Variable Energy Resources

This paper proposes an improved model for the reliability assessment of power systems considering component failures of variable energy resources (VER). The inherent intermittency of VER such as solar photovoltaic (PV) and wind farms, along with their susceptibility to component failures, present significant challenges to reliable system operation. These issues, combined with power grid operation and network constraints, complicate the reliable operation of VER-integrated power systems. Here, to address these concerns, this paper introduces a reliability assessment framework that considers VER input variability, its impact on component availability, and their resulting impact on overall system reliability. Stochastic models based on discrete Markov processes are developed to incorporate variable irradiance, wind speeds, and their effects on PV and wind component failure rates. A next-event and state transition-based approach is then developed to integrate the stochastic models into a mixed-timing sequential Monte Carlo simulation framework for composite reliability assessment. Case studies on the RTS-GMLC system demonstrate the effectiveness of the proposed model in evaluating the reliability of VER-integrated systems.

Pandit, Dilip [Sandia National Laboratories (SNL-N↗

The Corrective Maintenance Paradigm Shift at Hanford's Tank Farms - 20077

Hanford's Tank Farms facilities have been used to safely store waste for over 70 years, with the first single-shell tanks being constructed in 1943. Tank Farm facilities consist of 149 single-shell tanks, 28 double-shell tanks, an evaporator facility, and wastewater treatment facilities. Tank Farm facilities are aging, with a tremendous corrective maintenance burden on the Tank Farm contractor. The mission of Tank Farm facilities is soon changing from waste storage to waste staging for the Hanford Waste Treatment and Immobilization Plant (WTP). WTP operations will demand a significant increase in Tank Farm facility operations, in which corrective maintenance outage windows will shrink drastically. This realization has forced the Tank Farm contractor to consider a paradigm shift in Tank Farm facilities Maintenance planning, and the use of reliability Engineering tools. The Tank Farm Production Operations Engineering Cognizant System Engineering (CSE) organization has led the way in motivating this paradigm shift. This shift has been realized through the use of: 1) technical exchange with other Department of Energy (DOE) contractors to develop improvements in the CSE program, 2) a shift from the use of lagging to leading system health indicators, and 3) a Plant Health Committee to unite Engineering, Operations, and Maintenance personnel toward a productive maintenance strategy. The CSE organization has held several technical exchanges with other DOE contractors to discuss CSE concepts, and how to better maintain aging infrastructure. The technical exchange with other contractors has greatly reduced the time required to make improvements in the Tank Farm CSE program. Other DOE contractors have already faced issues surrounding aging infrastructure, and have vast experience in improving the reliability and usable life of structures and components in nuclear facilities. The past CSE program used lagging health indicators to determine the health of systems. The key lagging indicator used to determine system health was availability, which is the percentage of time that a facility was ready for operation compared to the time the facility was demanded for operation. Availability was a good indicator of health in the waste storage mission of Tank Farms, where safe storage was the most important function of the facility, and where maintenance outage windows were typically long-duration. In current and future operations, outage windows are reducing, resulting in the need for much more reliable systems. Systems that have had high availability may suddenly become inoperable due to a failed component or sub-system. In several instances, the use of availability as an indicator of system health failed to predict system/equipment failure before its occurrence. In discussions with other DOE contractors, a set of reliability tools, including leading indicators of health, has been implemented in the CSE program. This primarily involves the use of failure modes and effects analysis and the study of equipment failure to develop system monitoring plans that focus on trending data to detect oncoming equipment failure ahead of time. In addition, the use of a Plant Health Committee has added significantly to the paradigm shift from a corrective maintenance philosophy to the use of predictive and preventive maintenance. The Plant Health Committee is a chartered team consisting of Engineering, Operations, and Maintenance personnel. CSEs use this forum to present the results of their performance monitoring, including the presentation of health via leading health indicators. The most positive aspect of this committee is the communication that it creates within these critical organizations. The Operations and Maintenance organization benefit from focusing maintenance on the reliability-centered focus provided by Engineering. Engineering benefits from the operational experience of the Operations organization and from the failure data that can be provided by Maintenance personnel. The continued use of the Plant Health Committee is expected to further decrease maintenance outage times, in better support of oncoming 24/7 operations. (authors)

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

Nuclear Safety [Vol. 30, No. 3, July-September 1989]

Nuclear Safety is a review journal that covers significant developments in the field of nuclear safety. Its scope includes the analysis and control of hazards associated with nuclear energy, operations involving fissionable materials, and the products of nuclear fission and their effects on the environment. Primary emphasis is on safety in reactor design, construction, and operation; however, the safety aspects of the entire fuel cycle, including fuel fabrication, spent-fuel processing, nuclear waste disposal, handling of radioisotopes, and environmental effects of these operations, are also treated. Table of Contents for this issue follows. GENERAL SAFETY CONSIDERATIONS: 325 Safety of Framatome Advanced Nuclear Steam Supply Systems Designs by J. A. Charles and D. Lange, 333 Book Review of Nuclear Accidents: Intervention Levels for the Protection of the Public by H. B. Piper; ACCIDENT ANALYSIS: 335 Living PRA Computer Systems by S. C. Dinsmore and H.-P. Balfanz, 343 Summary of ICAP Assessments of RELAP5/MOD2 by W. E. Driskell and R. G. Hanson; CONTROL AND INSTRUMENTATION: 352 Thermal Performance Monitoring System at Maanshan Nuclear Power Plant by H.-J. Chao, Y.-P. Lin, G.-H. Jou, L.-Y. Liao, and Y.-B. Chen; DESIGN FEATURES: 358 Warning Systems for Nuclear Power Plant Emergencies by J. H. Sorensen and D. S. Mileti; WASTE AND SPENT FUEL MANAGEMENT: 371 Activities Related to Waste Management Compiled by E G. Silver; OPERATING EXPERIENCES: 382 Steam Generator Tube Performance: Experience with Water-Cooled Nuclear Power Reactors During 1985 by O. S. Tatone and R. L. Tapping, 400 Systems Interaction Analyses: Concepts and Techniques (Part II) by M. D. Muhlheim and G. A. Murphy, 413 Reactor Shutdown Experience Compiled by J. W. Cletcher, 416 Operating U.S. Power Reactors Compiled by E G. Silver; RECENT DEVELOPMENTS: 440 General Administrative Activities Compiled by E G. Silver, 460 Reports, Standards, and Safety Guides by D. S. Queener, 466 Status of Power-Reactor Licensing Activities Compiled by E G. Silver, 470 Proposed Rule Changes as of Mar. 31, 1989; ANNOUNCEMENTS: 334 Proceedings Published, 351 CEC Seminar on Methods and Codes for Assessing the Off-Site Consequences of Nuclear Accidents, 357 Short Course on Multiphase Flow and Heat Transfer: Bases and Applications in A: The Nuclear Power Industry B: The Process Industries, 357 International Conference on Probabilistic Safety Assessment and Management, 478 International Topical Meeting on the Safety, Status, and Future of Non-Commercial Reactors and Irradiation Facilities, 475 The Authors.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Sequence Diagrams & PFMEA Table - VGI [SWR-25-107]

As part of the VGI work under the National Charging Experience (ChargeX) Consortium, reliability analysis of communication interfaces for multiple SCM/VGI use-cases was performed using a Process Failure Modes and Analysis (PFMEA) style framework. This repository hosts all the relevant files for each of these use-cases which include: *A visual representation of their communication architecture: Image file (.png) *UML sequence diagram: Plant-UML source file (.puml). Visio file (.vsdx) and image file (.png) derived from the UML sourceX` *The PFMEA table: Excel file (.xlsx) These files are meant to serve as a starting point and can be adapted to company / organization specific SCM implementation.

Gadamsetty, Pranav [National Renewable Energy Labo↗

Characterization and identification of HPC applications at leadership computing facility

High Performance Computing (HPC) is an important method for scientific discovery via large-scale simulation, data analysis, or artificial intelligence. Leadership-class supercomputers are expensive, but essential to run large HPC applications. The Petascale era of supercomputers began in 2008, with the first machines achieving performance in excess of one petaflops, and with the advent of new supercomputers in 2021 (e.g., Aurora, Frontier), the Exascale era will soon begin. However, the high theoretical computing capability (i.e., peak FLOPS) of a machine is not the only meaningful target when designing a supercomputer, as the resources demand of applications varies. A deep understanding of the characterization of applications that run on a leadership supercomputer is one of the most important ways for planning its design, development and operation. In order to improve our understanding of HPC applications, user demands and resource usage characteristics, we perform correlative analysis of various logs for different subsystems of a leadership supercomputer. This analysis reveals surprising, sometimes counter-intuitive patterns, which, in some cases, conflicts with existing assumptions, and have important implications for future system designs as well as supercomputer operations. For example, our analysis shows that while the applications spend significant time on MPI, most applications spend very little time on file I/O. Combined analysis of hardware event logs and task failure logs show that the probability of a hardware FATAL event causing task failure is low. Combined analysis of control system logs and file I/O logs reveals that pure POSIX I/O is used more widely than higher level parallel I/O. Based on holistic insights of the application gained through combined and co-analysis of multiple logs from different perspectives and general intuition, we engineer features to "fingerprint" HPC applications. We use t-SNE (a machine learning technique for dimensionality reduction) to validate the explainability of our features and finally train machine learning models to identify HPC applications or group those with similar characteristic. To the best of our knowledge, this is the first work that combines logs on file I/O, computing, and inter-node communication for insightful analysis of HPC applications in production.

Liu, Zhengchun↗

Reliability Assessment of Cooling Fans for PV Inverters: Testing, Modeling, and Case Studies

The reliability of photovoltaic (PV) inverters is critical for long-term solar system performance, with cooling fan failures frequently leading to costly downtime. While much research exists on general cooling fan reliability, little attention has been given to fans operating within PV inverters and their unique environmental challenges. Here, this article proposes a comprehensive methodology to address this gap. First, a failure mode and effects analysis is performed on fans to identify the key failure mechanisms in PV applications, their corresponding stressors, and the models necessary for lifetime prediction. Second, an accelerated life test is designed and conducted to collect valuable experimental data for PV inverter fans in a reasonable amount of time. Third, a mathematical conversion of dynamic mission profiles into effective constant stress levels is derived. Fourth, case studies are given, showcasing lifetime estimates that account for geographic variations in mission profile data. The results demonstrate that this integrated approach leads to an accurate reliability assessment for PV inverter cooling fans.

accelerated life testing (ALT)↗