Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Reliability assessment”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Assembly and Integration Status of a High Fidelity Ground Test Bed for the Water Processor Assembly

The Water Recovery System (WRS) is a critical component of life support aboard the International Space Station (ISS) and will play an essential role in future missions beyond Low Earth Orbit (LEO). Its primary functional units – the Urine Processor Assembly (UPA), Brine Processor Assembly (BPA), and Water Processor Assembly (WPA) – must be evaluated for extended operation, dormancy resilience, material obsolescence, and reliability under exploration-driven constraints. Ground testing is vital for developing these technologies and generating statistically relevant reliability assessments, which requires extended runtime under integrated, Flight-like conditions. Currently, no high-fidelity, fully integrated WPA ground test bed exists to support these objectives. To address this gap, NASA is developing a WPA test bed at Marshall Space Flight Center (MSFC) that combines downgraded ISS flight hardware with functionally flight-like components in a cost-effective configuration while maintaining priority hardware investigations. This paper describes the current status of hardware assembly and integration, outlines key challenges such as simulating microgravity effects and mitigating obsolescence, and presents future test objectives including software development, reliability assessments, dormancy studies, and exploration-oriented upgrades.

Mary-Elizabeth Davis↗

A cost assessment of reliability requirements for shuttle-recoverable experiments

The relaunching of unsuccessful experiments or satellites will become a real option with the advent of the space shuttle. An examination was made of the cost effectiveness of relaxing reliability requirements for experiment hardware by allowing more than one flight of an experiment in the event of its failure. Any desired overall reliability or probability of mission success can be acquired by launching an experiment with less reliability two or more times if necessary. Although this procedure leads to uncertainty in total cost projections, because the number of flights is not known in advance, a considerable cost reduction can sometimes be achieved. In cases where reflight costs are low relative to the experiment's cost, three flights with overall reliability 0.9 can be made for less than half the cost of one flight with a reliability of 0.9. An example typical of shuttle payload cost projections is cited where three low reliability flights would cost less than $50 million and a single high reliability flight would cost over $100 million. The ratio of reflight cost to experiment cost is varied and its effect on the range in total cost is observed. An optimum design reliability selection criterion to minimize expected cost is proposed, and a simple graphical method of determining this reliability is demonstrated.

Campbell, J. W.↗

Study of SEM induced current and voltage contrast modes to assess semiconductor reliability

The purpose of the scanning electron microscopy study was to review the failure history of existing integrated circuit technologies to identify predominant failure mechanisms, and to evaluate the feasibility of their detection using SEM application techniques. The study investigated the effects of E-beam irradiation damage and contamination deposition rates; developed the necessary methods for applying the techniques to the detection of latent defects and weaknesses in integrated circuits; and made recommendations for applying the techniques.

Beall, J. R.↗

Assessing the Reliability of NDE

Versatile FORTRAN computer algorithm developed for calculating and plotting reliability of nondestructive evaluation (NDE) technique for inspection of flaws. Developed specifically to determine reliability of radiographic and ultrasonic methods for detection of critical flaws in structural ceramic materials. Reliability displayed in form of plot of probability of detection versus flaw size. NDE methods used in such applications as diagnostic medicine, quality control in industrial production, and prediction of failure in structural components.

Roth, Don J.↗

Pivotal-Function Assessment Of Reliability Of Software

Approach developed to establish utility of pivotal functions for estimation and prediction of reliability of software. Improved estimates of reliability with statistical confidence obtained when relatively few testing data available. Pivotal functions effective tools for determination of confidence limits for reliability of software and prediction limits for time to next failure. Provides exact confidence and prediction limits regardless of how many bugs found in software.

Hayhurst, Kelly J.↗

Using probabilistic analysis to assess the reliability of predicted SRB aft-skirt stresses

Probabilistic failure analysis is a tool to predict the reliability of a part or system. Probabalistic techniques were used to predict critical stresses which occur in the solid rocket booster aft-skirt during main engine buildup, immediately prior to lift-off. More than any other hold down post (HDP) load component, the Z loads are sensitive to variations in strains and calibration constants. Also, predicted aft-skirt stresses are strongly affected by HDP load variations. Therefore, the instrumented HDP are not effective load transducers for Z loads, and, when used with aft skirt stress indicator equations, yield estimates with large uncertainty. Monte Carlo simulation proved to be a straight forward way of studying the overlapping effects of multiple parameters on predicted equipment performance. An advantage of probabilistic analysis is the degree of uncertainty of each parameter as stated explicitly by its probability distribution. It was noted, however, that the choice of parameter distribution had a large effect on the simulation results. Many times these distributions must be assumed. The engineer who is designing the part should be responsible for the choice of parameter distribution.

Richardson, James A.↗

Issues and Methods for Assessing COTS Reliability, Maintainability, and Availability

Many vendors produce products that are not domain specific (e.g., network server) and have limited functionality (e.g., mobile phone). In contrast, many customers of COTS develop systems that am domain specific (e.g., target tracking system) and have great variability in functionality (e.g., corporate information system). This discussion takes the viewpoint of how the customer can ensure the quality of COTS components. In evaluating the benefits and costs of using COTS, we must consider the environment in which COTS will operate. Thus we must distinguish between using a non-mission critical application like a spreadsheet program to produce a budget and a mission critical application like military strategic and tactical operations. Whereas customers will tolerate an occasional bug in the former, zero tolerance is the rule in the latter. We emphasize the latter because this is the arena where there are major unresolved problems in the application of COTS. Furthermore, COTS components may be embedded in the larger customer system. We refer to these as embedded systems. These components must be reliable, maintainable, and available, and must be with the larger system in order for the customer to benefit from the advertised advantages of lower development and maintenance costs. Interestingly, when the claims of COTS advantages are closely examined, one finds that to a great extent these COTS components consist of hardware and office products, not mission critical software [1]. Obviously, COTS components are different from custom components with respect to one or more of the following attributes: source, development paradigm, safety, reliability, maintainability, availability, security, and other attributes. However, the important question is whether they should be treated differently when deciding to deploy them for operational use; we suggest the answer is no. We use reliability as an example to justify our answer. In order to demonstrate its reliability, a COTS component must pass the same reliability evaluations as the custom components, otherwise the COTS components will be the weakest link in the chain of components and will be the determinant of software system reliability. The challenge is that there will be less information available for evaluating COTS components than for custom components but this does not mean we should despair and do nothing. Actually, there is a lot we can do even in the absence of documentation on COTS components because the customer will have information about how COTS components are to be used in the larger system. To illustrate our approach, we will consider the reliability, maintainability, and availability (RMA) of COTS components as used in larger systems. Finally, COTS suppliers might consider increasing visibility into their products to assist customers in determining the components' fitness for use in a particular application. We offer ideas of information that would be useful to customers, and what vendors might do to provide it.

Schneidewind, Norman F.↗

Pre-Proposal Assessment of Reliability for Spacecraft Docking with Limited Information

This paper addresses the problem of estimating the reliability of a critical system function as well as its impact on the system reliability when limited information is available. The approach addresses the basic function reliability, and then the impact of multiple attempts to accomplish the function. The dependence of subsequent attempts on prior failure to accomplish the function is also addressed. The autonomous docking of two spacecraft was the specific example that generated the inquiry, and the resultant impact on total reliability generated substantial interest in presenting the results due to the relative insensitivity of overall performance to basic function reliability and moderate degradation given sufficient attempts to try and accomplish the required goal. The application of the methodology allows proper emphasis on the characteristics that can be estimated with some knowledge, and to insulate the integrity of the design from those characteristics that can't be properly estimated with any rational value of uncertainty. The nature of NASA's missions contains a great deal of uncertainty due to the pursuit of new science or operations. This approach can be applied to any function where multiple attempts at success, with or without degradation, are allowed.

Brall, Aron↗

Bionutrients-1: Utilizing Genomics and Transcriptomics to Assess the Reliability of Microorganisms for In Situ Nutrient Production on Long Duration Missions

The resupply of current long-duration crewed missions to the ISS relies on ground-launched supplies. As NASA looks toward Mars, ground-based resupply will no longer be an option. Critical nutrients, including vitamin C, vitamin K, folate, and thiamin, degrade during long-term storage, and regular consumption of these nutrients is essential for astronaut health. Another challenge of current food systems is the difficulty of consuming sufficient calories when subsisting on the limited flavors of freeze-dried food, which can lead to weight loss. The inclusion of microorganism-based food systems could alleviate both concerns. For example, the fermentation of rehydrated milk into yogurt with microorganisms genetically incorporating genes to produce critical vitamins would allow for both in situ production of nutrients and a fresh food product with additional flavor profiles. In comparison to plant food production, microorganisms require less flight infrastructure. The BioNutrients-1 mission is demonstrating viability of microbial fermentation food production in microgravity and testing the reliability of this approach for long-duration missions lacking resupply. While the BioNutrients-1 mission includes the collection of multiple phenotypic measurements, this status update will focus on the processing of samples for genomics and transcriptomics analyses as well as the planned analysis pipelines. First, the BioNutrients-1 mission seeks to identify microorganisms capable of surviving long-duration storage at ambient temperatures while maintaining genetic fidelity. To achieve this, nine commonly employed microbial species were stored at ambient temperatures in Stasis Packs on the ISS for five years. The viability and mutation rates will be measured at multiple time points for both flown and ground control samples. From an omics perspective, the changes in the bulk rates of point mutations and genetic rearrangements across the Stasis Pack species during the five years of storage will be determined, providing valuable insights into the potential of these microorganisms for long-duration space missions. Second, the BioNutrients-1 mission is characterizing the impact of microgravity on fermentation. Two strains of the yeast Saccharomyces cerevisiae, each encoding antioxidants (β-carotene or zeaxanthin) were flown to ISS for storage and fermentation within simplified bioreactors (Production Packs). The impact of microgravity on the expression of the antioxidant production genes and general metabolic genes will be determined using RNA sequencing. Ultimately, the transcriptome data will be compared to phenotypic measurements, such as the antioxidant yield, end-state biomass, and the production of EtOH, to determine the impacts of microgravity and long-term storage on microbial fermentation. The findings from this research will be instrumental in understanding the challenges and opportunities of microorganism-based food systems in space missions.

BioNutrients↗

Integrated performance and reliability specification for digital avionics systems

This paper describes an automated tool for performance and reliability assessment of digital avionics systems, called the Automated Design Tool Set (ADTS). ADTS is based on an integrated approach to design assessment that unifies traditional performance and reliability views of system designs, and that addresses interdependencies between performance and reliability behavior via exchange of parameters and result between mathematical models of each type. A multi-layer tool set architecture has been developed for ADTS that separates the concerns of system specification, model generation, and model solution. Performance and reliability models are generated automatically as a function of candidate system designs, and model results are expressed within the system specification. The layered approach helps deal with the inherent complexity of the design assessment process, and preserves long-term flexibility to accommodate a wide range of models and solution techniques within the tool set structure. ADTS research and development to date has focused on development of a language for specification of system designs as a basis for performance and reliability evaluation. A model generation and solution framework has also been developed for ADTS, that will ultimately encompass an integrated set of analytic and simulated based techniques for performance, reliability, and combined design assessment.

Brehm, Eric W.↗

Reliability modeling of fault-tolerant computer based systems

Digital fault-tolerant computer-based systems have become commonplace in military and commercial avionics. These systems hold the promise of increased availability, reliability, and maintainability over conventional analog-based systems through the application of replicated digital computers arranged in fault-tolerant configurations. Three tightly coupled factors of paramount importance, ultimately determining the viability of these systems, are reliability, safety, and profitability. Reliability, the major driver affects virtually every aspect of design, packaging, and field operations, and eventually produces profit for commercial applications or increased national security. However, the utilization of digital computer systems makes the task of producing credible reliability assessment a formidable one for the reliability engineer. The root of the problem lies in the digital computer's unique adaptability to changing requirements, computational power, and ability to test itself efficiently. Addressed here are the nuances of modeling the reliability of systems with large state sizes, in the Markov sense, which result from systems based on replicated redundant hardware and to discuss the modeling of factors which can reduce reliability without concomitant depletion of hardware. Advanced fault-handling models are described and methods of acquiring and measuring parameters for these models are delineated.

Bavuso, Salvatore J.↗

A Prognostic Launch Vehicle Probability of Failure Assessment Methodology for Conceptual Systems Predicated on Human Causal Factors

Create an improved method to calculate reliability of a conceptual launch vehicle system prior to fabrication by using historic data of actual root causes of failures. While failures have unique "proximate causes", there are typically a finite amount of common "root causes". Heretofore launch vehicle reliability evaluation typically hardware-centric statistical analyses, while most root causes of failures are been shown to be human-centric. A method based on human-centric root causes can be used to quantify reliability assessments and focus proposed actions to mitigate problems. Existing methods have been optimistic in their projections of launch vehicle reliability compared to actuals. Hypothesis: reliability of a conceptual launch vehicle can be more accurately evaluated based on a rational, probabilistic approach using past failure assessment teams' findings predicated on human-centric causes."Human Reliability Analysis Methods Selection Guidance for NASA"Chandler F.T., et al., NASA HQ/OSMA study group, July 2006. Outside HRA experts from academia, other federal labs, and the private sector. 50 system reliability methods considered, fourteen selected for further study, four finally selected as best suited for human spaceflight. Probabilistic Risk Analysis (PRA) + Human Reliability Analysis (HRA) enabled incorporating effects and probabilities of human errors. While four down-selected methods deemed appropriate for failure assessment, it did not appear that these methods could be concisely applied to perform major system-wide assessment of probability of failure of a conceptual design without becoming unwieldy."Engineering a Safer World", Detailed, comprehensive study external to NASA Leveson N. G., MIT, 2011.Systems-Theoretic Accident Model and Processes (STAMP). All-encompassing accident model based on systems theory analyzed accidents after they occurred and created approaches to prevent occurrence in developing systems not focused on failure prevention per se, but rather reducing hazards by influencing human behavior through use of constraints, hierarchical control structures, and process models to improve system safetySystem Theoretic Process Analysis (STPA) addresses predictive part of problem (a "hazard analysis"). Includes all causal factors identified in STAMP: "...design errors, software flaws, component interaction accidents, cognitively complex human decision-making errors, and social organizational and management factors contributing to accidents" can guide design process rather than require it to exist before-hand did not appear capable of concise application for system-wide assessment of probability of failure of a conceptual design without becoming unwieldy.

Williams, Craig H.↗

Reliability Studies for Fatigue-Crack Detection

Reusable test panels available to assess reliability of techniques that use fluorescent penetrant to detect fatigue cracks. Ultrasonic cleaning method developed for removing penetrant from panels prior to reuse.

Christner, B. K.↗

LOFT Debriefings: An Analysis of Instructor Techniques and Crew Participation

This study analyzes techniques instructors use to facilitate crew analysis and evaluation of their Line-Oriented Flight Training (LOFT) performance. A rating instrument called the Debriefing Assessment Battery (DAB) was developed which enables raters to reliably assess instructor facilitation techniques and characterize crew participation. Thirty-six debriefing sessions conducted at five U.S. airlines were analyzed to determine the nature of instructor facilitation and crew participation. Ratings obtained using the DAB corresponded closely with descriptive measures of instructor and crew performance. The data provide empirical evidence that facilitation can be an effective tool for increasing the depth of crew participation and self-analysis of CRM performance. Instructor facilitation skill varied dramatically, suggesting a need for more concrete hands-on training in facilitation techniques. Crews were responsive but fell short of actively leading their own debriefings. Ways to improve debriefing effectiveness are suggested.

Dismukes, R. Key↗

HRA Aerospace Challenges

Compared to equipment designed to perform the same function over and over, humans are just not as reliable. Computers and machines perform the same action in the same way repeatedly getting the same result, unless equipment fails or a human interferes. Humans who are supposed to perform the same actions repeatedly often perform them incorrectly due to a variety of issues including: stress, fatigue, illness, lack of training, distraction, acting at the wrong time, not acting when they should, not following procedures, misinterpreting information or inattention to detail. Why not use robots and automatic controls exclusively if human error is so common? In an emergency or off normal situation that the computer, robotic element, or automatic control system is not designed to respond to, the result is failure unless a human can intervene. The human in the loop may be more likely to cause an error, but is also more likely to catch the error and correct it. When it comes to unexpected situations, or performing multiple tasks outside the defined mission parameters, humans are the only viable alternative. Human Reliability Assessments (HRA) identifies ways to improve human performance and reliability and can lead to improvements in systems designed to interact with humans. Understanding the context of the situation that can lead to human errors, which include taking the wrong action, no action or making bad decisions provides additional information to mitigate risks. With improved human reliability comes reduced risk for the overall operation or project.

DeMott, Diana↗

CARE 3 User's Workshop

A user's workshop for CARE 3, a reliability assessment tool designed and developed especially for the evaluation of high reliability fault tolerant digital systems, was held at NASA Langley Research Center on October 6 to 7, 1987. The main purpose of the workshop was to assess the evolutionary status of CARE 3. The activities of the workshop are documented and papers are included by user's of CARE 3 and NASA. Features and limitations of CARE 3 and comparisons to other tools are presented. The conclusions to a workshop questionaire are also discussed.

Source record↗