Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Failure Rate”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 145 records · Page 8

Design for Reliability (DfR) in Space Life Support

The engineering process of Design for Reliability (DfR) is well established in the automotive and aerospace industries. DfR should be useful in the future development of space life support systems. DfR is a sequence of tasks that develop system requirements and plan reliability analysis and testing. First and fundamentally, the reliability requirement is defined. Next the system reliability model is developed, often using a reliability block diagram. The overall system reliability requirement is allocated to the subsystems and an estimate of the attainable reliability is made. This expected reliability can be improved by simplifying the design by removing components or by replacing less reliable components. Improving reliability can require difficult compromises, such as reducing performance requirements, increasing budget, or extending testing. The actual system reliability can be determined only by testing, which should continue long enough to provide the required confidence in the measured value. New systems often have unexpected design errors that cause failures in early testing. The usual reliability improvement process of testing, finding the failure modes, and redesigning to remove them reduces the failure rate and is referred to as “reliability growth.” After redesign has been completed, the system should be further tested to determine the actual achieved reliability more accurately. If the final system failure rate is too high, redundant systems can be used to improve overall operational reliability. Adding redundancy simply to increase the one- or two-fault tolerance metric may sometimes reduce reliability. Reliability can be improved in three ways: redesigning the system to include more reliable subsystems and components, reliability growth testing and failure mode removal, and by using parallel redundant systems. DfR should combine these approaches to achieve the required reliability while managing performance, cost, and schedule.

Reliability↗

Surf-Deformer: Mitigating Dynamic Defects on Surface Code via Adaptive Deformation

In this paper, we introduce Surf-Deformer, a code deformation framework that seamlessly integrates adaptive defect mitigation functionality into the current surface code workflow. It crafts several basic deformation instructions based on fundamental gauge transformations, which can be combined to explore a larger design space than previous methods. This enables more optimized deformation processes tailored to specific defect situations, restoring the QEC capability of deformed codes more efficiently with minimal qubit resources. Additionally, we design an adaptive code layout that accommodates our defect mitigation strategy while ensuring efficient execution of logical operations. Our evaluation shows that Surf-Deformer outperforms previous methods by significantly reducing the end-to-end failure rate of various quantum programs by 35× to 70×, while requiring only about 50% of the qubit resources compared to the previous method to achieve the same level of failure rate. Ablation studies show that Surf-Deformer surpasses previous defect removal methods in preserving QEC capability and facilitates surface code communication by achieving nearly optimal throughput.

Yin, Keyi↗

Risk and Performance Assessment of Generic Mission Architectures: Showcasing the Artemis Mission

A has initiated a strong push to return face. In this work, we astronaut assess performance and risk for proposed mission architectures using a new Mission Architecture Risk Assessment (MARA) tool. The MARA tool can produce statistics about the availability of components and overall performance of the mission considering potential failures of any of its components. In a Monte Carlo approach, the tool repeats the mission simulation multiple times while a random generator lets modules fail according to their failure rates. The results provide statistically meaningful insights into the overall performance of the chosen architecture. A given mission architecture can be freely replicated in the tool, with the mission timeline and basic characteristics of employed mission modules (habitats, rovers, power generation units, etc.) specified in a configuration file. Crucially, failure rates for each module need to be known or estimated. The tool performs an event-driven simulation of the mission and accounts for random failure events. Failed modules can be repaired, which takes crew time but restores operations. In addition to tracking individual modules, MARA can assess the availability of predefined functions throughout the mission. For instance, the function of resource collection would require a rover to collect the resources, a power generation unit to charge the rover, and a resource processing module. Together, the modules that are required for a given function are called a functional group. Similarly, we can assess how much crew time is available to achieve a mission benefit (e.g. research, building a base, etc) as opposed to spending crew time on repairs. Here we employ the method on the proposed NASA Artemis mission. Artemis aims to return United States astronauts to the lunar surface by 2024. Results provide insights into mission failure probabilities, up- and downtime for individual modules and crew-time resources spent on the repair of failed modules. The tool also allows us to tweak the mission architecture in order to find setups that produce more favorable mission performance. As such, the tool can be an aid in improving the mission architect abling cost-benefit analysis for mission improvement.

Rumpf, Clemens M.↗

Metal Bellows Valve Reliability Testing - Copper Stem Tip Testing

SRNL was funded in Mid-Year FY20 by NNSA NA-231 to continue evaluation of alternate valves for use in tritium service to support domestic Mo-99 production. The focus of the effort was to identify valve cycle life as a function of actuator size and stem tip material. Using the minimum size actuator to reliability open and close valves can reduce glovebox size and thus support domestic companies to “come to market” faster in supplying Mo-99 to the US market. This report serves as a continuation to the FY19 report and summarizes the task activities completed in FY20 after authorization to start work was obtained on May 5 th , 2020. Copper stem tips were tested on the Swagelok 1C and 5C actuated metal bellows valves. With ambitions to cycle each set 150,000 times, both 1C and 5C valves were met with high failure rates. The smaller 1C actuated valves required additional closing pressure to form a seal with the Cu stem tips installed however, excessive stem tip deformation may be the root cause of the majority of the valves failing before 500 cycles. The larger 5C actuated valves were cycled 150,000 times but still resulted in 80% failure rate, suspected of metal fatigue in the bellows due to high cycling frequencies.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Dispenser Reliability: Materials R&D. A Hydrogen Fueling Infrastructure Research and Station Technology (H2FIRST) Report

Dispensers are the top cause of maintenance events and down-time at hydrogen fueling stations. In an effort to help characterize and enable improvements in dispenser reliability, an extensive accelerated lifetime testing set-up was designed and built at NREL involving components typically part of dispensing operations at fueling stations. Device Under Test (DUTs) included different components such as normally open valves, normally closed valves, fueling nozzles, breakaways devices and filters. Conditions of testing included pressures, and flow rates similar to light duty fuel cell electric vehicles fueling at -40°C, and -20°C for thousands of cycles in hydrogen. Tested components (failed and non-failed) were disassembled at SNL and polymeric O-rings were carefully retrieved and cataloged for chemical and physical characterization. Data collected was compared to similar O-rings from unexposed or non-tested components for hydrogen effects, and failure modes. Degradation analyses, based on select polymer chemistries common across all component types, their location within components, visual assessment of damage coupled with strong hydrogen effects from chemical characterization, was completed and presented to NREL and DOE. Overall, the failure rate amongst the components was not as high as expected for the test conditions. Among the component types tested, breakaways were the most susceptible to damage under these test conditions, with fueling nozzles a close second. The proper combination of selection of the right polymer and optimum component design was found to make a strong difference in component reliability under severe dispenser operating conditions. Physical degradation of polymers, rather than chemical changes due to low temperature hydrogen exposure, is more prevalent as failure mode for these test conditions. The nature and the extent of the degradation was much less at -20°C as compared to -40°C. The damage and failure rates were higher at lower temperatures than at higher test temperatures. As expected, increasing the number of cycles at the lowest test temperature (-40°C) increased damage. This indicates that cycling at the low temperature of -40°C required by SAE J2601 can reduce component life in fuel dispensing operations

08 HYDROGEN↗

Impacts of PV Module Connector Failures on Cost and Performance of Utility Scale Photovoltaic Systems

The reliability, cost and performance of electrical connectors are a concern in all types of electrical systems, and demands on connectors used on photovoltaic (PV) systems include that connectors maintain electrical conductivity and physical strength, endure ultraviolet sunlight and high ambient temperature, and resist moisture and chemical intrusion over a very long (>25 year) performance period. Connector failures increase operation and maintenance (O&M) costs and reduce plant production, but connector failure can also cause safety and liability problems, which are of greater concern. This work results from a three-year collaboration between Sandia National Laboratories (SNL), the Electric Power Research Institute (EPRI), and the National Renewable Energy Laboratory (NREL) and funded by the U.S. Department of Energy (DOE) Solar Energy Technology Office (SETO) under Agreements #39035 and #38531 "Connector Reliability Across the US Solar Sector." a multi-pronged investigation of PV connector health across the US (see https://energy.sandia.gov/pvconnectors/). This report presents derivation of a Techno-Economic Analysis (TEA) that models failure modes and frequencies (how often failure occurs), estimates O&M costs and lost production associated with connector failures, and then calculates the effect that PV module connectors can have on Levelized Cost of Energy (LCOE). The model is informed with initial data from quantitative assessment of failure rates, root causes and mechanisms, in-situ diagnostics and data collection, lab-based forensics, and interviews with PV connector manufacturers and plant operators. SNL conducted site inspections at multiple utility-scale sites in different climates and subjected field samples of new, used, and degraded connectors to visual and electrical characterization. EPRI conducted metallurgical analysis of the pin and sleeve conductors to study failure-induced morphological and compositional changes. There is in general a shortage of statistically valid data, but data from PVROM database maintained by SNL was sufficient to ascertain failure rates and lost production as well as provide qualitative insight in its curated maintenance records. This report details the structure of the mathematical model but the sources of data to inform the model will continue to evolve. Analysis of a 100 MW PV plant is provided as an example of the use of the model, with results indicating that connectors are responsible for Annualized O&M Costs of $\$$71,933/year; Annualized Unit O&M Costs of $\$$0.72/kW/year; that a Reserve Account of $\$$187,220 should be available to fund repairs related to connectors; that connectors add $\$$1,494,004 to the Net Present Value of the O&M Costs (project life); and that O&M related to connectors adds about $\$$0.00088/kWh to the Levelized Cost of Energy. The impact of this model is to provide a tool to make the US solar sector more robust by quantifying and monetizing the reliability risks to utility-scale PV systems posed by poorly installed, mismatched and/or poorly designed and manufactured connectors. The TEA provides a model incorporating failure statistics, O&M cost data, and lost production into a single figure of merit, informing decisions and enabling practitioners to optimize cost and performance trade-offs. Stakeholders include connector manufacturers, system designers and equipment specifiers, standards bodies, installers and O&M providers, investors and insurance underwriters. This report supports continued growth of PV predicated on assurances that properly installed and maintained PV system connectors are safe and reliable. The project team is proposing future work including accelerated testing of connectors and expanding the approach taken here to other PV system components, such as TEA for rapid shut-down devices.

14 SOLAR ENERGY↗

NASA Helps Keep the Light Burning for the Saturn Car Company

The Saturn Electronics & Engineering, Inc. (Saturn) facility in Marks, Miss., that produces lamp assemblies was experiencing itermittent problems with its automotive under the hood lamps. After numerous testing and engineering efforts, technicians could not pin down the root of the problem. So Saturn contacted the NASA Technology Assistance Program (TAP) at Stennis Space Center. The Marks production facility had been experiencing intermittent problems with under the hood lamp assemblies for some time. The failure rate, at 2 percent, was unacceptable. Every effort was made to identify the problem so that corrective action could be put in place. The problem was investigated and researched by Saturn's engineering department. In addition, Saturn brought in several independent testing laboratories. Other measures included examining the switch component suppliers and auditing them for compliance to the design specifications and for surface contaminants. All attempts to identify the factors responsible for the failures were inconclusive. In an effort to get to the root of the problem, and at the recommendation of the Mississippi Department of Economic Development, Saturn contacted the NASA TAP at Stennis. The NASA Materials and Contamination Laboratory, with assistance from the Stennis Prototype Laboratory, conducted a materials evaluation study on the switch components. The laboratory findings showed the failures were caused by a build-up of carbon-based contaminants on the switch components. Saturn Electronics & Engineering, Inc., is a minority-owned provider of contract manufacturing services to a diverse global marketplace. Saturn operates manufacturing facilities globally serving the North American, European, and Asian markets. Saturn's production facility in Marks, Mississippi, produces more than 1,000,000 lamps and switches monthly. "Since the NASA recommendations were implemented, our internal failure rate for intermittency has dropped to less than .02 percent. Most importantly, we restored our high-level of customer satisfaction. Stennis provided an invaluable service to our business," Patrick said. Both NASA and Saturn were pleased with the results form this technical assistance project. The Technology Assistance Program at Stennis makes available to the public NASA technical expertise and access to lab facilities. This project provided both services with a positive outcome.

Source record↗

Using Technical Performance Measures

All programs have requirements. For these requirements to be met, there must be a means of measurement. A Technical Performance Measure (TPM) is defined to produce a measured quantity that can be compared to the requirement. In practice, the TPM is often expressed as a maximum or minimum and a goal. Example TPMs for a rocket program are: vacuum or sea level specific impulse (lsp), weight, reliability (often expressed as a failure rate), schedule, operability (turn-around time), design and development cost, production cost, and operating cost. Program status is evaluated by comparing the TPMs against specified values of the requirements. During the program many design decisions are made and most of them affect some or all of the TPMs. Often, the same design decision changes some TPMs favorably while affecting other TPMs unfavorably. The problem then becomes how to compare the effects of a design decision on different TPMs. How much failure rate is one second of specific impulse worth? How many days of schedule is one pound of weight worth? In other words, how to compare dissimilar quantities in order to trade and manage the TPMs to meet all requirements. One method that has been used successfully and has a mathematical basis is Utility Analysis. Utility Analysis enables quantitative comparison among dissimilar attributes. It uses a mathematical model that maps decision maker preferences over the tradeable range of each attribute. It is capable of modeling both independent and dependent attributes. Utility Analysis is well supported in the literature on Decision Theory. It has been used at Pratt & Whitney Rocketdyne for internal programs and for contracted work such as the J-2X rocket engine program. This paper describes the construction of TPMs and describes Utility Analysis. It then discusses the use of TPMs in design trades and to manage margin during a program using Utility Analysis.

Garrett, Christopher J.↗

Efficient and Reliable Power Takeoff for Ocean Wave Energy Harvesting

The project goal is to significantly improve the current ocean wave energy harvesting through innovative Power Take-off (PTO) design, advanced power electronics, and novel wave capture structures. The objective of the project is to design and demonstrate system-agnostic components for application across multiple MHK systems, and complete component designs, build scaled prototypes, and perform testing and analysis for metric validation of 25% increase in component rating/per unit cost and 50% reduction in failure rate. The major innovation of the PTO is the Mechanical Motion Rectifier (MMR) mechanism that rectifies the bi-directional oscillatory motion of the input from waves into a steady unidirectional rotation output to directly drive the electrical generator. Through this mechanism, the efficiency and the fatigue life of the PTO can be significantly improved to benefit the energy absorption and lifespan of the wave energy converters (WEC). During the period of performance, the component and system design are completed, the scaled prototypes are developed and performed testing. It is validated that the 25% increase in a component rating/per unit cost. The 50% reduction in failure rate is not directly validated by experiments, however, it can be explained qualitatively with analysis. Besides, 8 journal articles, 13 conference proceedings, 1 patent, 3 Master thesis and two Ph.D. dissertations are published based on the work related to this project. The list of all the publications can be found at the end of the project as an appendix. Over 30 students and postdocs were trained through this project. Three prototypes of 100W and 500W WECs and 10KW PTO were designed, built, and tested in ocean wave tank and using the NREL dynamometer. This project demonstrated 50-80% PTO efficiency, 90-98% power electronics efficiency, up to 66% capture width ratio in irregular waves, and 34% overall efficiency in regular waves.

16 TIDAL AND WAVE POWER↗

Management of Risks Associated with Application of Novel Materials in Novel Operating Environments in Novel Reactor Designs

There is currently no widely agreed, detailed general method for licensing a novel plant incorporating novel materials (or materials being deployed in novel environments); in many such situations, there are no directly applicable engineering code cases for decision-makers (including regulators) to rely on. This paper discusses a framework for solving this problem that is based on the Reliability and Integrity Management (RIM) approach delineated in ASME BPVC Section XI Division 2. NRC Regulatory Guide 1.246, Rev. 0, endorses, with conditions, the subject portion of the ASME Code. The proposed framework is meant to support development of a licensing case by addressing certain technical challenges. The framework discussed here is compatible with the Licensing Modernization Project, but applying it in a specific case will call for advances in the state of practice, if not the state of the art. The RIM approach calls for applicants to (a) allocate reliability targets to plant structures, systems, and components (SSCs), (b) show that they are able to relate the currently observed physical condition of each SSC in the program to its failure probability well enough to determine whether the target reliability allocations are being satisfied, allowing for uncertainty related to the novelty of the materials/designs/operating environments, and (c) be able to demonstrate that the proposed program of surveillances will reliably detect unacceptable degradation of an SSC before SSC failure occurs. These challenges are discussed in the paper, and a potentially applicable modeling approach based on cumulative damage rather than failure rates is briefly illustrated.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Integrated Risk-Informed Condition Based Maintenance Capability and Automated Platform: Technical Report 1

Due to continuing global energy market trends, driven heavily by the abundant preserves of natural gas, there is an immediate need to reduce costs associated with operation and maintenance (O&M) for the current domestic nuclear power industry and for future reactor developments. This is to ensure that nuclear power generation remains an economically competitive and viable option in the energy market. O&M costs include labor-intensive preventive maintenance (PM) programs, which involve manually-performed inspection, calibration, testing, and maintenance of plant assets at periodic frequency and time-based replacement of assets, irrespective of their condition. This has resulted in an expensive, labor-centric business model to achieve high capacity factors. Fortunately, there are technologies (advanced sensors, data analytics, and risk assessment methodologies) that can enable the transition from a labor-centric business model to a technology-centric business model. The technology-centric business model will result in a significant reduction of PM activities, laying the foundation for real-time condition assessment of plant assets, reducing overall labor and part costs. To enable this transition, PKMJ Technical Services LLC is partnering with the U.S. Department of Energy’s Idaho National Laboratory (operated by the Battelle Energy Alliance, LLC) and the Public Services Enterprise Group (PSEG) Nuclear, LLC in the Integrated Risk-Informed Condition-Based Maintenance Capability and Automated Platform Project. In this report, the configuration of a digital cloud platform using Microsoft Azure is discussed, data from the PSEG Salem Nuclear Generating Station Units 1 & 2 are imported into a digital cloud platform, and the data is used for an evaluation of several key areas: cost impact analysis, risk-informed model development, and preventive maintenance strategy optimization. First, the cost impact analysis reviews which plant assets are potential good candidates for condition-based monitoring. Next, INL utilized the data in their local environment to develop the risk-informed model; which provides estimates of failure rates and probability of failures of assets based upon their past performance. The developed model is performed on assets selected from the cost impact analysis. Lastly, engineers assess the preventive maintenance strategy for the selected assets at PSEG against maintenance strategies in the nuclear industry for similar assets to potentially identify acceptable justification for the extension of current maintenance frequencies.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

JANTX1N5420 diode

Testing of sample lots from Unitrode and Micro Semiconductor had to be stopped in group 1 test because 50% failure rate limit was reached. Failure analysis was performed only for group 2 testing because of apparent failure mode.

Source record↗

General Monte Carlo reliability simulation code including common mode failures and HARP fault/error-handling

A Monte Carlo Fortran computer program was developed that uses two variance reduction techniques for computing system reliability applicable to solving very large highly reliable fault-tolerant systems. The program is consistent with the hybrid automated reliability predictor (HARP) code which employs behavioral decomposition and complex fault-error handling models. This new capability is called MC-HARP which efficiently solves reliability models with non-constant failures rates (Weibull). Common mode failure modeling is also a specialty.

Platt, M. E.↗

Modeling Predictive Maintenance for NuScale’s Condensate and Feedwater System Using EMRALD

Event Modeling Risk Assessment using Linked Diagrams (EMRALD) is used to model of NuScale’s feedwater and condenser system to capture component degradation and repairing process. To model the reliability of the components, a three-stage failure rate is used contrary to a constant rate leading to component failure. Results from EMRALD will be used to compare different maintenance strategies to reduce system downtime. The outcome of the project will assist in maximizing remaining useful life of the overall system and increasing the availability and revenue of the plant.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Study to evaluate the effect of EVA on payload systems. Volume 1: Executive summary

Programmatic benefits to payloads are examined which can result from the routine use of extravehicular activity (EVA) during space missions. Design and operations costs were compared for 13 representative baseline payloads to the costs of those payloads adapted for EVA operations. The EVA-oriented concepts developed in the study were derived from these baseline concepts and maintained mission and program objectives as well as basic configurations. This permitted isolation of cost saving factors associated specifically with incorporation of EVA in a variety of payload designs and operations. The study results were extrapolated to a total of 74 payload programs. Using appropriate complexity and learning factors, net EVA savings were extrapolated to over $551M for NASA and U.S. civil payloads for routine operations. Adding DOD and ESRO payloads increases the net estimated savings of $776M. Planned maintenance by EVA indicated an estimated $168M savings due to elimination of automated service equipment. Contingency problems of payloads were also analyzed to establish expected failure rates for shuttle payloads. The failure information resulted in an estimated potential for EVA savings of $1.9 B.

Patrick, J. W.↗

Field failure mechanisms for photovoltaic modules

Beginning in 1976, Department of Energy field centers have installed and monitored a number of field tests and application experiments using current state-of-the-art photovoltaic modules. On-site observations of module physical and electrical degradation, together with in-depth laboratory analysis of failed modules, permits an overall assessment of the nature and causes of early field failures. Data on failure rates are presented, and key failure mechanisms are analyzed with respect to origin, effect, and prospects for correction. It is concluded that all failure modes identified to date are avoidable or controllable through sound design and production practices.

Dumas, L. N.↗

Resilient U.S. Land Ports of Entry

The continued operation of Land Ports of Entry (LPOE), managed by the Customs and Border Protection (CBP) and General Services Administration% is vital to the U.S. economy and security. Border faculties are included in the Department of Homeland Security (DHS) Government Facilities Sector2, one of the 16 critical infrastructures "whose assets, systems, and networks, whether physical or virtual, are considered so vital to the United States that their incapacitation or destruction would have a debilitating effect on security, national economic security, national public health or safety, or any combination thereof.'" Specifically, disruptions to the flow of border crossing traffic, in the form of closures or increased border crossing wait times, impact the economy and security of all countries involved. This paper describes a process for analyzing and improving the resilience of U.S. Land Ports of Entry. For LPOE, the team believes that energy resilience is the primary objective due to the complete reliance on the e-manifest system and the increasing use of Multi-Energy Portals (MEPs). Emanifests are part of CPB's Automated Commercial Environment (ACE). They document several key pieces of information about cargo vehicles wishing to cross the border into the United States and are submitted before arriving at the port. Vehicles can be flagged for more invasive inspection based on the content of the e-manifest. MEPs are a non-intrusive inspection (NII) technology used to scan the contents of the cargo. Together MEPs and ACE serve an important role in aiding CBP with their mission to protect "the public from dangerous people and materials", and "enabling legitimate trade and travel.'" To analyze resilience of a port, the team would need to understand the port's current energy usage, which systems depend on energy and what backup systems exist, and any emergency operation plans that dictate how systems are operated in the event of a power outage. The team would also need to determine the design basis threats (DBTs) for the LPOE which could include natural disasters, manmade events, and accidents. The magnitudes of the DBTs are calculated and are then translated to expected impacts on the infrastructure and systems at the port. With this information gathered, existing LPOE models developed here at Sandia National Laboratories could be extended to support decisions about resilience. Current models are implemented in FlexSim, a 3rd party discrete event simulator. FlexSim provides 3-D visuals of physical layout that can reveal valuable insights, allows input to be variable (e.g. time it takes to interact with the CBP officer at primary inspection can vary) so that a whole range of possibilities can be captured in the results, and can be used to collect user-defined output metrics. Current LPOE models focus on cargo vehicle traffic, and process changes caused by the installation of new drive-through MEPs. Extending them to address resilience questions would require the addition of key pieces of information learned during the resilience analysis including critical systems, failure rates, and process changes for when failures occur. The primary output metric for current models is border crossing wait time. Additional metrics would also be added to the model to gain a more complete understanding of impacts related to resilience, for example, MEP scan rate. Once complete, the model could be used to analyze the effectiveness of mitigation strategies representing some future state.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗