Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Reliability and Integrity Management”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Prototype thin-film thermocouple/heat-flux sensor for a ceramic-insulated diesel engine

A platinum versus platinum-13 percent rhodium thin-film thermocouple/heat-flux sensor was devised and tested in the harsh, high-temperature environment of a ceramic-insulated, low-heat-rejection diesel engine. The sensor probe assembly was developed to provide experimental validation of heat transfer and thermal analysis methodologies applicable to the insulated diesel engine concept. The thin-film thermocouple configuration was chosen to approximate an uninterrupted chamber surface and provide a 1-D heat-flux path through the probe body. The engine test was conducted by Purdue University for Integral Technologies, Inc., under a DOE-funded contract managed by NASA Lewis Research Center. The thin-film sensor performed reliably during 6 to 10 hr of repeated engine runs at indicated mean surface temperatures up to 950 K. However, the sensor suffered partial loss of adhesion in the thin-film thermocouple junction area following maximum cyclic temperature excursions to greater than 1150 K.

Kim, Walter S.↗

Hybrid Power Management

An engineering discipline denoted as hybrid power management (HPM) has emerged from continuing efforts to increase energy efficiency and reliability of hybrid power systems. HPM is oriented toward integration of diverse electric energy-generating, energy-storing, and energy-consuming devices in optimal configurations for both terrestrial and outer-space applications. The basic concepts of HPM are potentially applicable at power levels ranging from nanowatts to megawatts. Potential applications include terrestrial power-generation, terrestrial transportation, biotechnology, and outer-space power systems. Instances of this discipline at prior stages of development were reported (though not explicitly labeled as HPM) in three prior NASA Tech Briefs articles: "Ultracapacitors Store Energy in a Hybrid Electric Vehicle"(LEW-16876), Vol. 24, No. 4 (April 2000), page 63; "Photovoltaic Power Station With Ultracapacitors for Storage" (LEW-17177), Vol. 27, No. 8 (August 2003), page 38; and "Flasher Powered by Photovoltaic Cells and Ultracapacitors" (LEW-17246), Vol. 24, No. 10 (October 2003), page 37. As the titles of the cited articles indicate, the use of ultracapacitors as energy-storage devices lies at the heart of HPM. An ultracapacitor is an electrochemical energy-storage device, but unlike in a conventional rechargeable electrochemical cell or battery, chemical reactions do not take place during operation. Instead, energy is stored electrostatically at an electrode/electrolyte interface. The capacitance per unit volume of an ultracapacitor is much greater than that of a conventional capacitor because its electrodes have much greater surface area per unit volume and the separation between the electrodes is much smaller. Power-control circuits for ultracapacitors can be simpler than those for batteries, for two reasons: (1) Because of the absence of chemical reactions, charge and discharge currents can be greater than those in batteries, limited only by the electrical resistances of conductors; and (2) whereas the charge level of a battery depends on voltage, temperature, age, and load condition, the charge level of an ultracapacitor, like that of a conventional capacitor, depends only on voltage.

Eichenberg, Dennis↗

Bridging Equipment Reliability Data and Risk Informed Decisions in a Plant Operation Context

Industry equipment reliability and asset management programs are essential elements that help ensure the safe and economical operation of nuclear power plants. The effectiveness of these programs is addressed in several industry-developed and regulatory programs. The Risk-Informed Asset Management (RIAM) project is tasked to develop tools in support of the equipment reliability and asset management programs at nuclear power plants. These tools are designed to create a direct bridge between component health/lifecycle data and decision making (e.g., maintenance scheduling and project prioritization). The goal of this article is to provide a guide for specific use cases that the RIAM project is targeting. We have grouped uses cases into three main areas. The first area focuses on the analysis of equipment reliability data with a particular emphasis on condition-based data, such as test/surveillance reports and component monitoring data. The second area focuses on the integration of equipment reliability into system/plant reliability models to determine system/plant health and identify the components that are critical to maintain an operational system. Lastly, the third area manages plant resources, such as maintenance activities and replacement scheduling using optimization methods. Here the primary focus is on supporting typical system engineer decisions regarding maintenance activity scheduling and component aging management. This is performed in a risk-informed context where the term “risk” is broadly constructed to include both plant reliability and economics. This framework combines data analytics tools to analyze equipment reliability data with risk-informed methods designed to support system engineer decisions (e.g., maintenance and replacement schedules, optimal maintenance posture) in a customizable workflow.

97 - MATHEMATICS AND COMPUTING↗

Space station software reliability analysis based on failures observed during testing at the multisystem integration facility

Quality of software not only is vital to the successful operation of the space station, it is also an important factor in establishing testing requirements, time needed for software verification and integration as well as launching schedules for the space station. Defense of management decisions can be greatly strengthened by combining engineering judgments with statistical analysis. Unlike hardware, software has the characteristics of no wearout and costly redundancies, thus making traditional statistical analysis not suitable in evaluating reliability of software. A statistical model was developed to provide a representation of the number as well as types of failures occur during software testing and verification. From this model, quantitative measure of software reliability based on failure history during testing are derived. Criteria to terminate testing based on reliability objectives and methods to estimate the expected number of fixings required are also presented.

Tamayo, Tak Chai↗

SIGHT: Stacked Integration of Geospatial Hierarchical Typologies for Inferring Building Characteristics

Building characteristics are often absent in building stock datasets, particularly in regions most vulnerable to climate change and requiring effective disaster management strategies. Traditional machine learning approaches, while widely used to predict building attributes, typically neglect the spatial context of the data, leading to less accurate and reliable outcomes. To address these challenges, this paper introduces a novel algorithm, the Stacked Integration of Geospatial Hierarchical Typologies. This algorithm adapts a meta-learning framework to incorporate geospatial context into the predictive modeling process. We demonstrate the utility of the algorithm through two primary use cases: building use type classification and building height prediction. The algorithm consistently achieved or exceeded a 0.94 macro average F1 score across five geographically distinct countries for building use type classification. For building height prediction, it accurately predicted heights with a root mean square error of 3.01 in a comprehensive study using roughly 3.6 million buildings in Japan. These results underscore the benefits of integrating spatial hierarchies into machine learning models, enhancing both predictive accuracy and reliability in geospatial modeling. This work introduces a new algorithm to address the pervasive data sparsity issue in existing building stock datasets.

Adams, Daniel [ORNL] (ORCID:0000000196950577)↗

The evolution of the Human Systems and Simulation Laboratory in nuclear power research

The events at Three Mile Island in the United States brought about fundamental changes in the ways that simulation would be used in nuclear operations. The need for research simulators was identified to scientifically study human-centered risk and make recommendations for process control system designs. This paper documents the human factors research conducted at the Human Systems and Simulation Laboratory (HSSL) since its inception in 2010 at Idaho National Laboratory. The facility’s primary purposes are to provide support to utilities for system upgrades and to validate modernized control room concepts. In the last decade, however, as nuclear industry needs have evolved, so too have the purposes of the HSSL. Thus, beyond control room modernization, human factors researchers have evaluated the security of nuclear infrastructure from cyber adversaries and evaluated human-in-the-loop simulations for joint operations with an integrated hydrogen generation plant. Lastly, our review presents research using human reliability analysis techniques with data collected from HSSL-based studies and concludes with potential future directions for the HSSL, including severe accident management and advanced control room technologies.

21 SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLANTS↗

Enabling Ultra-Compact, Lightweight, Efficient, and Reliable 6.6 kW On-Board Bi-Directional Electric Vehicle Charger with Advanced Topology and Control

The research explored new topologies, control methods, mechanical integration, and thermal management methods for electric vehicle (EV) on-board chargers. The team investigated capacitor-based power conversion, leveraging the high energy densities inherent to capacitive energy storage compared to inductive methods. The proposed topologies simultaneously enabled high power density and high efficiency of the design. The proposed architecture was demonstrated in a 6.6 kW bi-directional charger prototype. The research pursued several directions to improve system performance. Innovative topologies were studied for both the main power conversion stage as well as the single-phase twice-line-frequency energy buffer. To ensure robust and efficient operation, new control methods were developed to integrate these two subsystems. To achieve high power density in the full system solution, the mechanical structure of the charger is highly optimized to maximally fill the converter box volume. In parallel with the mechanical design effort, the converter was packaged with high-performance cooling methods which removed heat from key areas of power dissipation in the converter. The thermal management system was optimized to minimize its weight and volume, ultimately motivating the design of a custom additively manufactured cold-plate. The full system achieves a peak power of 7 kW with less than 0.3% total harmonic distortion (THD) and greater than 0.994 power factor in power factor correction (PFC) operation, corresponding to a total box-volume power density of 47.9 kW/L and gravimetric power density of 24.6 W/g. The system achieves a peak efficiency of 98.9%, with 97.9% efficiency at maximum power.

33 ADVANCED PROPULSION SYSTEMS↗

Electric Access System Enhancement (EASE): Assessment of a Distributed Energy Resource Management System for Enabling Dynamic Hosting Capacity

The Electric Access System Enhancement (EASE) project demonstrated a scalable, interoperable, and cost-effective means of integrating high penetration of distributed energy resources (DERs). The control architecture developed leveraged a Distributed Energy Resource Management System, Distribution System Operator for transacting energy, and 3rd Party DER Aggregator platform. The project identified ways to enhance the customer interconnection process to the grid and improve access to information from DERs and optimize the usage of DERs to provide energy services in a simulated day-ahead shadow market while maintain grid reliability. By integrating these capabilities into a scalable system of systems, the DSO can effectively balance DER generation and customer demand on the distribution network. This capability allows the grid to host more DER than traditionally possible on wires alone. Hosting more DER has the added benefit of supplying increased demand growth, sometimes beyond the capacity limits of the distribution network itself. This is known as a dynamic hosting capacity, which could help utilities manage the forecasted growth in electricity demand as California switch to electric vehicles and appliances. This could help to establish energy storage and PV generation as a “pseudo firm” generation resource mix if managed appropriately and provide sufficient resource adequacy for distribution capacity upgrade deferrals.

14 SOLAR ENERGY↗

A Review of Quantum Computing Technologies in Power System Optimization

As modern power grids increasingly integrate variable renewable generation, distributed energy resources, and energy storage systems, classical optimization techniques are facing unprecedented challenges. This review examines the emerging application of quantum computing to overcome these challenges in power system optimization, including optimal power flow (OPF), unit commitment (UC), economic dispatch (ED), and intelligent switching and topology optimization (IS-TO). Recent research has introduced various quantum methodologies—such as gate-based, annealing-based, variational algorithms, and quantum-inspired algorithms—to address the combinatorial complexity inherent in grid reconfiguration and energy management. The review summaries the quantum algorithms, quantum devices and the power system test cases, highlighting hybrid quantum–classical strategies that leverage the complementary strengths of both paradigms. Some quantum advantages have been observed, including theoretical speedup, accurate simulation results, scalable qubit usage, efficient QUBO mapping. In particular, the review emphasizes the importance of integrating quantum optimization techniques with classical control frameworks, these hybrid approaches demonstrate the potential to improve real-time grid management and operational reliability. A significant portion of the analysis is devoted to the practical limitations of current quantum devices. Present-day quantum hardware, operating in the noisy intermediate-scale quantum (NISQ) era, remains highly sensitive to noise and limited in qubit connectivity, which constrains the scale and accuracy of implemented algorithms. The review delves into specific challenges such as the need for qubit-efficient encoding techniques and error mitigation strategies that are critical for handling real-world grid optimization problems. In addition, the work draws attention to the performance discrepancies between theoretical quantum speedups and experimental validations, underscoring the importance of rigorous benchmark studies using representative power grid test cases. In summary, this review highlights both the promise and limitations of quantum computing for power system optimization. It provides a comprehensive overview of the state-of-the-art technologies, categorizes recent advancements in algorithm design, and discusses practical considerations for implementation, and serves as an informative resource on current research. Future research directions include developing robust hybrid frameworks, advancing qubit-efficient formulations, and scaling up experimental demonstrations to confirm the theoretical advantages of quantum methods in large-scale power system operations.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Automated Inventory Solutions in End-User IT Support

Managing IT equipment by hand is prone to errors and delays, severely impacting operational continuity and productivity. Manual inventory systems often result in time delays, inconsistent record-keeping, equipment shortages, and increased workloads for IT staff. At Savannah River National Laboratory (SRNL), my internship focused on creating an automated inventory management solution using Microsoft Power Automate and SharePoint Lists. This solution seamlessly integrates with the existing Microsoft 365 infrastructure, thus eliminating the need for additional software purchases or dedicated server space. By providing real-time updates and reducing manual data entry, the new system ensures a more reliable and maintainable approach to IT asset management.

Information Technology↗

Laboratory testing methods to evaluate the reliability of occupancy sensors for commercial building applications

The energy performance of commercial buildings is greatly influenced by occupants which are highly variable and among the most unpredictable components of a building's operation. While most building control systems use fixed, predetermined occupancy schedules, these fixed occupancy levels can be quite different from actual occupancy. This can cause unnecessary energy consumption, particularly from heating, ventilation, and air conditioning (HVAC) and lighting systems which are responsible for approximately 60% of commercial buildings' energy use. The use of occupancy counting sensor systems integrated with building management system controls is one method that can be used to improve the energy-consuming performance of buildings. However, there is no standardized universal methodology and metrics to evaluate their reliability. The aim of this research is to develop a uniform evaluation methodology to assess the reliability of occupancy counting sensor systems in a controlled laboratory environment. The developed testing methodology includes both “typical” scenarios representing the occupancy scenarios of a typical commercial building, and “failure” testing scenarios which represent a range of potential scenarios that may impact a sensor system's reliability. These methods were then implemented in a case study to evaluate the performance of two novel occupancy counting sensor systems (i.e., door-centric, and camera-based). Results suggest that typical testing results can be used to compare the overall performance of the occupancy counting sensor systems; however, failure testing is also important to understand the weaknesses of the sensor system in order to select the suitable one for the intended use of the commercial building. In addition, the proposed methodology includes a modified confusion matrix which enables the ability to identify if failures are caused by over or under counting occupants and to what extent this occurs over the testing period.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

The NASA IVHM Technology Experiment for X-37

The NASA IVHM (Integrated Vehicle Health Management) technology experiment for X-37 is presented. The goals and objectives of this program are: to reduce cost and increase reliability of space transportation; to demonstrate benefits of in-flight IVHM to the operation of a Reusable Launch Vehicle; to advance this IVHM technology to Technology Readiness Level approx. 7 within a flight environment; and to operate IVHM software on the Vehicle Management Computer. The following sections are included: Background (X-37 & Livingstone), Livingstone model example from DS-1, Experiment overview, X-37 IVHM scope, Stanley interface to livingstone model, Right ruddervator actuator, Motor state diagram, inferred nominal state, and X-37 informed maintenance experiment.

Source record↗

Emerging technologies for V&V of ISHM software for space exploration

Systems1,2 required to exhibit high operational reliability often rely on some form of fault protection to recognize and respond to faults, preventing faults' escalation to catastrophic failures. Integrated System Health Management (ISHM) extends the functionality of fault protection to both scale to more complex systems (and systems of systems), and to maintain capability rather than just avert catastrophe. Forms of ISHM have been utilized to good effect in the maintenance phase of systems' total lifecycles (often referred to as 'condition-based mainte-nance'), but less so in a 'fault protection' role during actual operations. One of the impediments to such use lies in the challenges of verification, validation and certification of ISHM systems themselves. This paper makes the case that state-of-the-practice V&V and certification techniques will not suffice for emerging forms of ISHM systems; however, a number of maturing software engineering assurance technologies show particular promise for addressing these ISHM V&V challenges.

fault detection, isolation, and recovery↗

Smart Sensor Demonstration Payload

Sensors are a critical element to any monitoring, control, and evaluation processes such as those needed to support ground based testing for rocket engine test. Sensor applications involve tens to thousands of sensors; their reliable performance is critical to achieving overall system goals. Many figures of merit are used to describe and evaluate sensor characteristics; for example, sensitivity and linearity. In addition, sensor selection must satisfy many trade-offs among system engineering (SE) requirements to best integrate sensors into complex systems [1]. These SE trades include the familiar constraints of power, signal conditioning, cabling, reliability, and mass, and now include considerations such as spectrum allocation and interference for wireless sensors. Our group at NASA s John C. Stennis Space Center (SSC) works in the broad area of integrated systems health management (ISHM). Core ISHM technologies include smart and intelligent sensors, anomaly detection, root cause analysis, prognosis, and interfaces to operators and other system elements [2]. Sensor technologies are the base fabric that feed data and health information to higher layers. Cost-effective operation of the complement of test stands benefits from technologies and methodologies that contribute to reductions in labor costs, improvements in efficiency, reductions in turn-around times, improved reliability, and other measures. ISHM is an active area of development at SSC because it offers the potential to achieve many of those operational goals [3-5].

Schmalzel, John↗

PR100: Puerto Rico Grid Resilience and Transition to 100% Renewable Energy

The Puerto Rico Grid Resilience and Transitions to 100% Renewable Energy Study (PR100) comprehensively analyzes possible pathways for Puerto Rico to achieve its renewable energy goals while incorporating stakeholder perspectives and advancing energy resilience for all Puerto Ricans. PR100 is a wide-ranging and in-depth examination of Puerto Rico's energy system investment options. The findings are the culmination of two years of stakeholder engagement, scenario modeling, and impact analysis. The PR100 report and Implementation Roadmap contain a range of results and actions that reflect Puerto Rico's priorities around energy justice, resilience, and reliability. Led by the U.S. Department of Energy's Grid Deployment Office with funding from the Federal Emergency Management Agency, the PR100 study leveraged and integrated dozens of best-in-class models and in-depth analyses from researchers across six national laboratories: National Renewable Energy Laboratory (which led the study), along with Argonne National Laboratory, Lawrence Berkeley National Laboratory, Oak Ridge National Laboratory, Pacific Northwest National Laboratory, and Sandia National Laboratories (which conducted the study). For more information, please see the "PR100 Project Website" resource below.

Array↗

From Here to Autonomicity: Self-Managing Agents and the Biological Metaphors that Inspire Them

We seek inspiration for self-managing systems from (obviously, pre-existing) biological mechanisms. Autonomic Computing (AC), a self-managing systems initiative based on the biological metaphor of the autonomic nervous system, is increasingly gaining momentum as the way forward for integrating and designing reliable systems, while agent technologies have been identified as a key enabler for engineering autonomicity in systems. This paper looks at other biological metaphors such as reflex and healing, heart- beat monitors, pulse monitors and apoptosis for assisting in the realization of autonomicity.

Sterritt, Roy↗

A rose by any other name: Certification seen as process rather than content

Green (1990) believes that the two main factors safeguarding flying from human error are both related to certification and regulation. First is the increasingly proceduralized nature of flying whereby as much as possible is reduced to a rule-based activity. Second is the emphasis placed upon training and competency checking of aircrew in simulators and in the air, both generally and for all particular types of aircraft flown. This leaves, believes Green, other human factors that are relatively unaddressed as yet and which can give rise to human reliability problems. These include: hardware factors and especially pilot/co-pilot relationships; and system factors including fatigue and cost/safety trade-offs. He also, importantly, identifies problems with the integration of the 'electronic crew member' following increased automation. Human reliability failures with artificial intelligence and automation, due to over-reliance on the system fail-safe mechanisms, or to operator under- confidence in the integrity or self-regulating capacity of the system, or to out-of-loop effects, are widely accepted as being due to deficiencies in plant design, planning, management and maintenance more than to 'operator error' - Reason's (1990) latent error or organization pathogens argument. Reliability failures in complex systems are well enough documented to give cause for concern and at least promote a debate on the merits of a full certification program. The purpose of this short paper is to seek out and explore what is valuable in certification, at the least to show that the benefits outweigh the disadvantages and at best to identify positive outcomes perhaps not obtainable in other ways. On both sides of the debate on certification there is general agreement on the need for a better human factors perspective and effort in complex aviation systems design. What is at issue is how this is to be promoted. It is incumbent upon opponents of certification to say how else such promotion be enabled. This is an exploratory and philosophical review, not a focused and specific one, and it will draw upon much that is not firmly in the domain of complex aviation systems.

Wilson, John R.↗

Modeling in the State Flow Environment to Support Launch Vehicle Verification Testing for Mission and Fault Management Algorithms in the NASA Space Launch System

Analysis methods and testing processes are essential activities in the engineering development and verification of the National Aeronautics and Space Administration's (NASA) new Space Launch System (SLS). Central to mission success is reliable verification of the Mission and Fault Management (M&FM) algorithms for the SLS launch vehicle (LV) flight software. This is particularly difficult because M&FM algorithms integrate and operate LV subsystems, which consist of diverse forms of hardware and software themselves, with equally diverse integration from the engineering disciplines of LV subsystems. M&FM operation of SLS requires a changing mix of LV automation. During pre-launch the LV is primarily operated by the Kennedy Space Center (KSC) Ground Systems Development and Operations (GSDO) organization with some LV automation of time-critical functions, and much more autonomous LV operations during ascent that have crucial interactions with the Orion crew capsule, its astronauts, and with mission controllers at the Johnson Space Center. M&FM algorithms must perform all nominal mission commanding via the flight computer to control LV states from pre-launch through disposal and also address failure conditions by initiating autonomous or commanded aborts (crew capsule escape from the failing LV), redundancy management of failing subsystems and components, and safing actions to reduce or prevent threats to ground systems and crew. To address the criticality of the verification testing of these algorithms, the NASA M&FM team has utilized the State Flow environment6 (SFE) with its existing Vehicle Management End-to-End Testbed (VMET) platform which also hosts vendor-supplied physics-based LV subsystem models. The human-derived M&FM algorithms are designed and vetted in Integrated Development Teams composed of design and development disciplines such as Systems Engineering, Flight Software (FSW), Safety and Mission Assurance (S&MA) and major subsystems and vehicle elements such as Main Propulsion Systems (MPS), boosters, avionics, Guidance, Navigation, and Control (GN&C), Thrust Vector Control (TVC), liquid engines, and the astronaut crew office. Since the algorithms are realized using model-based engineering (MBE) methods from a hybrid of the Unified Modeling Language (UML) and Systems Modeling Language (SysML), SFE methods are a natural fit to provide an in depth analysis of the interactive behavior of these algorithms with the SLS LV subsystem models. For this, the M&FM algorithms and the SLS LV subsystem models are modeled using constructs provided by Matlab which also enables modeling of the accompanying interfaces providing greater flexibility for integrated testing and analysis, which helps forecast expected behavior in forward VMET integrated testing activities. In VMET, the M&FM algorithms are prototyped and implemented using the same C++ programming language and similar state machine architectural concepts used by the FSW group. Due to the interactive complexity of the algorithms, VMET testing thus far has verified all the individual M&FM subsystem algorithms with select subsystem vendor models but is steadily progressing to assessing the interactive behavior of these algorithms with LV subsystems, as represented by subsystem models. The novel SFE applications has proven to be useful for quick look analysis into early integrated system behavior and assessment of the M&FM algorithms with the modeled LV subsystems. This early MBE analysis generates vital insight into the integrated system behaviors, algorithm sensitivities, design issues, and has aided in the debugging of the M&FM algorithms well before full testing can begin in more expensive, higher fidelity but more arduous environments such as VMET, FSW testing, and the Systems Integration Lab7 (SIL). SFE has exhibited both expected and unexpected behaviors in nominal and off nominal test cases prior to full VMET testing. In many findings, these behavioral characteristics were used to correct the M&FM algorithms, enable better test coverage, and develop more effective test cases for each of the LV subsystems. This has improved the fidelity of testing and planning for the next generation of M&FM algorithms as the SLS program evolves from non-crewed to crewed flight, impacting subsystem configurations and the M&FM algorithms that control them. SFE analysis has improved robustness and reliability of the M&FM algorithms by revealing implementation errors and documentation inconsistencies. It is also improving planning efficiency for future VMET testing of the M&FM algorithms hosted in the LV flight computers, further reducing risk for the SLS launch infrastructure, the SLS LV, and most importantly the crew.

Trevino, Luis↗