Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Reliability and Integrity Management”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 289 records · Page 16

Wireless Subsurface Sensors for Health Monitoring of Thermal Protection Systems on Hypersonic Vehicles

Health diagnostics is an area where major improvements have been identified for potential implementation into the design of new reusable launch vehicles (RLVs) in order to reduce life cycle costs, to increase safety margins, and to improve mission reliability. NASA Ames is leading the effort to develop inspection and health management technologies for thermal protection systems. This paper summarizes a joint project between NASA Ames and industry partners to develop "wireless" devices that can be embedded in the thermal protection system to monitor temperature or other quantities of interest. These devices are sensors integrated with radio-frequency identification (RFID) microchips to enable non-contact communication of sensor data to an external reader that may be a hand-held scanner or a large portal. Both passive and active prototype devices have been developed. The passive device uses a thermal fuse to indicate the occurrence of excessive temperature. This device has a diameter under 0.13 cm. (suitable for placement in gaps between ceramic TPS tiles on an RLV) and can withstand 370 C for 15 minutes. The active device contains a small battery to provide power to a thermocouple for recording a temperature history during flight. The bulk of the device must be placed beneath the TPS for protection from high temperature, but the thermocouple can be placed in a hot location such as near the external surface.

Milos, Frank S.↗

Pitfalls and Precautions When Using Predicted Failure Data for Quantitative Analysis of Safety Risk for Human Rated Launch Vehicles

Launch vehicle reliability analysis is largely dependent upon using predicted failure rates from data sources such as MIL-HDBK-217F. Reliability prediction methodologies based on component data do not take into account system integration risks such as those attributable to manufacturing and assembly. These sources often dominate component level risk. While consequence of failure is often understood, using predicted values in a risk model to estimate the probability of occurrence may underestimate the actual risk. Managers and decision makers use the probability of occurrence to influence the determination whether to accept the risk or require a design modification. The actual risk threshold for acceptance may not be fully understood due to the absence of system level test data or operational data. This paper will establish a method and approach to identify the pitfalls and precautions of accepting risk based solely upon predicted failure data. This approach will provide a set of guidelines that may be useful to arrive at a more realistic quantification of risk prior to acceptance by a program.

Hatfield, Glen S.↗

A Novel Architecture for Attack-Resilient Wide-Area Protection and Control System in Smart Grid

Wide-area protection and control (WAPAC) systems are widely applied in the energy management system (EMS) that rely on a wide-area communication network to maintain system stability, security, and reliability. As technology and grid infrastructure evolve to develop more advanced WAPAC applications, however, so do the attack surfaces in the grid infrastructure. This paper presents an attack-resilient system (ARS) for the WAPAC cybersecurity by seamlessly integrating the network intrusion detection system (NIDS) with intrusion mitigation and prevention system (IMPS). In particular, the proposed NIDS utilizes signature and behavior-based rules to detect attack reconnaissance, communication failure, and data integrity attacks. Further, the proposed IMPS applies state transition-based mitigation and prevention strategies to quickly restore the normal grid operation after cyberattacks. As a proof of concept, we validate the proposed generic architecture of ARS by performing experimental case study for wide-area protection scheme (WAPS), one of the critical WAPAC applications, and evaluate the proposed NIDS and IMPS components of ARS in a cyber-physical testbed environment. Our experimental results reveal a promising performance in detecting and mitigating different classes of cyberattacks while supporting an alert visualization dashboard to provide an accurate situational awareness in real-time.

24 POWER TRANSMISSION AND DISTRIBUTION↗

funcX: Federated Function as a Service for Science

Here, funcX is a distributed function as a service (FaaS) platform that enables flexible, scalable, and high performance remote function execution. Unlike centralized FaaS systems, funcX decouples the cloud-hosted management functionality from the edge-hosted execution functionality. funcX's endpoint software can be deployed, by users or administrators, on arbitrary laptops, clouds, clusters, and supercomputers, in effect turning them into function serving systems. funcX's cloud-hosted service provides a single location for registering, sharing, and managing both functions and endpoints. It allows for transparent, secure, and reliable function execution across the federated ecosystem of endpoints-enabling users to route functions to endpoints based on specific needs. funcX uses containers (e.g., Docker, Singularity, and Shifter) to provide common execution environments across endpoints. funcX implements various container management strategies to execute functions with high performance and efficiency on diverse funcX endpoints. funcX also integrates with an in-memory data store and Globus for managing data that may span endpoints. We motivate the need for funcX, present our prototype design and implementation, and demonstrate, via experiments on two supercomputers, that funcX can scale to more than 130000 concurrent workers. We show that funcX's container warming-aware routing algorithm can reduce the completion time for 3,000 functions by up to 61% compared to a randomized algorithm and the in-memory data store can speed up data transfers by up to 3x compared to a shared file system.

97 MATHEMATICS AND COMPUTING↗

Thermal Management for FPGA Nodes in HPC Systems

The integration of FPGAs into large-scale computing systems is gaining attention. In these systems, real-time data handling for networking, tasks for scientific computing, and machine learning can be executed with customized datapaths on reconfigurable fabric within heterogeneous compute nodes. At the same time, thermal management, particularly battling the cooling cost and guaranteeing the reliability, is a continuing concern. The introduction of new heterogeneous components into HPC nodes only adds further complexities to thermal modeling and management. The thermal behavior of multi-FPGA systems deployed within large compute clusters is less explored. Here, we first show that the thermal behaviors of different FPGAs of the same generation can vary due to their physical locations in a rack and process variation, even though they are running the same tasks. We present a machine learning–based model to capture the thermal behavior of each individual FPGA in the cluster. We then propose two thermal management strategies guided by our thermal model. First, we mitigate thermal variation and hotspots across the cluster by proactive thermal-aware task placement. Under the tested system and benchmarks, we achieve up to 26.4° C and on average 13.3° C system temperature reduction with no performance penalty. Second, we utilize this thermal model to guide HLS parameter tuning at the task design stage to achieve improved thermal response after deployment.

97 MATHEMATICS AND COMPUTING↗

Avista’s Shared Energy Economy Model Pilot: A Techno-economic Assessment

As part of the second round of the Washington Clean Energy Fund, Avista Corp received a $3.5 million matching grant in support of a shared energy economy project to test the integration of energy assets–from rooftop solar and battery storage to building energy management systems–that can be shared and used for multiple purposes. The goal of this project is to demonstrate how both the customer and the utility can benefit from this shared energy economy model and demonstrate that the electric grid can become more reliable, efficient, resilient, and flexible. Pacific Northwest National Laboratory was engaged by the U.S. Department of Energy and the Washington State Department of Commerce to work with Avista in assessing the benefits of the shared energy economy model. This report documents the techno-economic assessment of the shared energy economy model, including the definition of use cases and applications, collection and preparation of data and input parameters, development of modeling and optimization methods, and case studies and analysis results.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

An Integrated Framework for Risk Assessment of Safety-related Digital Instrumentation and Control Systems in Nuclear Power Plants: Methodology Refinement and Exploration

This report documents activities performed by Idaho National Laboratory (INL) during Fiscal Year (FY) 2023 for the U.S. Department of Energy (DOE) Light Water Reactor Sustainability (LWRS) Program, Risk Informed Systems Analysis (RISA) Pathway, digital instrumentation and control (DI&C) risk assessment project. In FY 2019, the RISA Pathway initiated a project to develop a risk assessment strategy for delivering a technical basis to support effective, and secure DI&C technologies for digital upgrades/designs. A risk assessment-informed framework was proposed for this strategy, which aims to (1) provide a best-estimate, risk informed capability to quantitatively estimate the safety margin obtained from plant modernization, especially for safety-related DI&C systems, (2) support and supplement existing risk informed DI&C design guides by providing quantitative risk information and evidence, (3) offer a capability of design architecture evaluation of various DI&C systems, (4) assure the long-term safety and reliability of safety-related DI&C systems, and (5) reduce uncertainty in costs and support integration of DI&C systems in the plant. To achieve these technical goals, the LWRS-developed framework provides a means to address relevant technical issues by: (1) defining a risk informed analysis process for DI&C upgrade that integrates hazard analysis, reliability analysis, and consequence analysis, (2) applying risk informed tools to address common cause failures (CCFs) and quantify corresponding failure probabilities for DI&C technologies, particularly software CCFs, (3) evaluating the impact of digital failures at the component level, system level, and plant level, and (4) providing insights and suggestions on designs to manage the risks, thus to support the development and deployment of advanced DI&C technologies in nuclear power plants (NPPs). Adding diversity within a system or components is the primary means to eliminate and mitigate CCFs, but diversity also increases system complexity and may not address all sources of systematic failures. Optimization of diversity and redundancy applications for the safety-critical DI&C systems remains a challenge. To deal with the technical issues in addressing potential software CCFs in safety-related DI&C systems of NPPs and supporting relevant design optimization, the proposed framework provides: (a) A best-estimate, risk informed capability to address new technical digital issues quantitatively, focusing on software CCFs in safety-related DI&C systems of NPPs; (b) A common and a modularized platform for DI&C designers, software developers, cybersecurity analysts, and plant engineers to predict and prevent risk in the early design stage of DI&C systems; (c) Technical bases and risk informed insights to assist users address the risk informed alternatives for evaluation of CCFs in safety-related DI&C systems of NPPs; and (d) A risk informed tool that offers a capability of design architecture evaluation of various DI&C systems to support system design decisions in diversity and redundancy applications. The research and development efforts of this project in FY 2023 are focused on refining current methods on software CCF modeling and estimation and exploring additional innovative approaches to risk assessment of DI&C systems to enable a more comprehensive and complete assessment of various safety-related DI&C design architectures. The primary audience of this report are DI&C designers, engineers, and probabilistic risk assessment (PRA) practitioners. This includes stakeholders, such as the nuclear utilities and regulators who consider the deployment and upgrade of DI&C systems, DI&C software developers and reviewers, and cybersecurity specialists. It should be noted that all the analyses are performed for the demonstration of the methodology, not for the evaluation of an actual digital control system. Results are obtained based on limited design information and testing data.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

A cost effective data management subsystem for the LST

The paper outlines the approach used in developing DMS (Data Management Subsystem) alternatives for the LST (Large Space Telescope) and in selecting the concept considered to be the most cost effective means of implementing the LST DMS requirements. Two candidate DMS concepts are discussed: a functionally integrated and a functionally separated one. For the single vehicle LST program, separation of the DMS functions best provides high reliability, operations flexibility, minimal interface complexity, and the least complex software development and verification task. The use of available hardware and NASA standard components is stressed.

Dougherty, J. A.↗

Effectiveness of Redundant Communications Systems in Maintaining Operational Control of Small Unmanned Aircraft

As a part of NASA’s Unmanned Aircraft System (UAS) Traffic Management (UTM) research, a test was performed to evaluate the effectiveness of the redundant Command and Control (C2) communications system for maintaining operational control of small UAS in the airspace over a rural area. In the test, operators set up a primary and a secondary UAS C2 communications system, sent a maneuver command to an Unmanned Aircraft (UA) with and without a functioning primary system, then verified the execution of the sent command to confirm the operator control. Operators reported that the tested redundancy configurations were effective in maintaining operational control in the test airspace over rural locations. Since the next phase of UTM research focuses on operations in an urban area where an increased level of Radio Frequency (RF) activities occur compared to a rural area, four recommendations are provided to sustain the effectiveness of redundancy in urban operations. First, the operator should not include C2 systems that use the industrial, scientific, and medical (ISM) radio bands in redundancy configurations. Second, the operator should verify the RF characteristics of the intended operation area and examine the area’s radio noise floor. Third, the operator should monitor the availability, quality, and reliability of communications services used by a redundant system. Fourth, the small UAS community should adopt a standard set of contingency steps to handle the loss of C2 communications so that such events are managed in a consistent manner across the airspace. The insights from the test will be used to accommodate the FAA’s UAS integration effort.

Jung, Jaewoo↗

Integration issues of a plasma contactor Power Electronics Unit

A hollow cathode-based plasma contactor is baselined on International Space Station Alpha (ISSA) for spacecraft charge control. The plasma contactor system consists of a hollow cathode assembly (HCA), a power electronics unit (PEU), and an expellant management unit (EMU). The plasma contactor has recently been required to operate in a cyclic mode to conserve xenon expellant and extend system life. Originally, a DC cathode heater converter was baselined for a continuous operation mode because only a few ignitions of the hollow cathode were expected. However, for cyclic operation, a DC heater supply can potentially result in hollow cathode heater component failure due to the DC electrostatic field. This can prevent the heater from attaining the proper cathode tip temperature for reliable ignition of the hollow cathode. To mitigate this problem, an AC cathode heater supply was therefore designed, fabricated, and installed into a modified PEU. The PEU was tested using resistive loads and then integrated with an engineering model hollow cathode to demonstrate stable steady-state operation. Integration issues such as the effect of line and load impedance on the output of the AC cathode heater supply and the characterization of the temperature profile of the heater under AC excitation were investigated.

Pinero, Luis R.↗

Adaptive Protection and Validated Models to Enable Deployment of High Penetrations of Solar PV (PV-MOD)

The availability and validation of various PV models in commercial tools differ, with some models not yet thoroughly validated for advanced inverter functionalities and reliable performance under weak system conditions. Many existing models do not fully incorporate new inverter control functions, which can affect system stability. The increasing deployment of solar PV and other inverter-based resources (IBRs), including distributed energy resources (DERs), is influencing the reliable operation of protection schemes in distribution systems and microgrids. Emerging adaptive protection schemes (APS) offer new opportunities for protecting these systems during varying configurations and DER operating conditions, though their demonstration and validation remain limited. Adaptive protection schemes face similar challenges, as they are typically designed for specific configurations. There is a growing need for tools and methodologies to streamline the deployment of adaptive protection for safe and reliable DER integration. The project main objective was to develop and validate high-fidelity generic models of solar PV facilities for stability, protection, EMT, and QSTS analyses. This objective was achieved, and these models can now be integrated into commercial software tools, enabling utilities, vendors, and developers to study high-penetration PV systems more confidently. The project also demonstrated advanced applications of these models, including the design and deployment of adaptive protection schemes in high-penetration field applications and microgrids, supporting grid safety and reliability. Several milestones were reached by the end of the project. A sophisticated inverter test plan was developed, and inverters representative of the North American marketplace were selected. EPRI and NREL tested various inverters, conforming to IEEE standards. Improvements were made to existing generic models of IBR units, IBR plants, and aggregated feeders for various analyses. The first generic electromagnetic transient (EMT) model for a solar PV plant was developed, conforming to IEEE Std 2800™-2022 and validated against laboratory measurements of a 2.2 MVA large-scale battery energy storage system (BESS) inverter. That model was then used to produce reference responses illustrating examples of validated and verified IBR plant models that pass or fail tests for technical minimum capability and performance as specified in the IEEE standard. The developed, tested, and validated generic models can be used for transmission planning, stability assessments, expansion planning, and evaluating potential future IBR interconnection requirements. They can also support interconnection screens and conformity assessments of IBR plants, including solar PV. The project significantly contributed to the ongoing standardization and model-based representation and verification of IBR responses. The project further addressed challenges of common distribution protection schemes with increasing deployment of DER by developing, validating, and demonstrating adaptive protection schemes (APS) that can improve the reliable and safe integration of DER into distribution systems. New APS were designed using improved DER models for three common distribution systems: a radial feeder, a meshed network, and a microgrid. Modeling and hardware-in-the-loop (HIL) testing of the APS were conducted, successfully showing their effectiveness and selectivity. Proof-of-concept field demonstration was achieved for two APS, i.e., one on a radial feeder and another one in a microgrid. Field demonstration could not be achieved for the APS on a meshed network, primarily due apprehension of one utility partner and also due to limited access to the protective algorithms in the network protectors. Guidelines developed from the lessons learned in the project lay out the general process followed in the design, installation, and commissioning of APS for various distribution systems. Distribution utility partners’ apprehension about field demonstration of the new APS were addressed—with varying success—by taking a stepped risk-management approach of modeling of a wide range of sensitivities first, performing in-depth proof-of-concept testing in the laboratory including HIL next, and finally deliberately implementing and commissioning the actual protection equipment and algorithms into parts of—or in parallel operation to—the three real distribution systems. Future work should include pilot projects that further show the acceptable performance of the developed APS before these schemes be rolled out more widely. Inclusion of both utility and original equipment manufacturers (OEMs) in future projects could increase chances of successful field demonstration. Despite challenges in achieving the field demonstration goal of the project for all three APS, the research significantly contributed to the innovation of adaptive protection solutions for scalable and reliable DER integration into distribution systems. This project significantly enhances the understanding of the impact of using appropriate inverter models on distribution and transmission (T&D) systems. By addressing the limitations of existing generic models, the project introduces high-fidelity models for stability, protection, electromagnetic transient (EMT), and quasi-static time series (QSTS) analyses. These models, integrated into commercial software tools, enable utilities, vendors, and developers to confidently study high-penetration PV systems. The project also demonstrates advanced applications, including adaptive protection schemes (APS) for distribution systems and microgrids, ensuring grid safety and reliability. The technical effectiveness and economic feasibility of the methods are evident through the development and validation of sophisticated inverter test plans and the selection of representative inverters. Testing by EPRI and NREL on retail, commercial, and utility-scale inverters, conforming to IEEE standards, underscores the robustness of the models. Improvements to existing generic models for various analyses further enhance their validity and applicability. The project also identifies gaps in common distribution protection schemes and designed new APS using improved DER models, demonstrating their effectiveness through modeling and hardware-in-the-loop (HIL) testing. The project’s benefits to the public are manifold. By advancing the standardization and model-based representation of IBR response, it supports transmission planning, stability assessments, and future IBR interconnection requirements. The generic models can facilitate better communication between transmission planners and developers, supporting expected IBR plant capability and performance. Additionally, the development of APS for radial feeders, meshed networks, and microgrids supports the integration of distributed energy resources (DERs) into distribution systems, enhancing grid reliability and safety. The project’s emphasis on thorough testing and simplicity in design ensures practical and scalable solutions for DER integration.

14 SOLAR ENERGY↗

Multi-Timescale Optimal Operation Framework for Integrated Economic and Reliability Analysis of Hybrid Power Plants

This paper introduces a hierarchical modeling framework for hybrid power plants (HPP) to facilitate the operation of HPP in power systems similar to conventional generators (Congens) in the integrated multi-timescale optimal operation framework. To consider the uncertainties of HPP renewable power in the day-ahead scheduling, distributionally robust optimization (DRO) is used. To ensure that the state-of-charge (SOC) of energy storage systems in HPPs aligns closely with the planned value for long-term reliability, real-time SOC management is incorporated. In addition, an adjustable real-time control is designed for the robust delivery of HPP real-time services. Case studies performed on a revised IEEE 39-bus system demonstrate the effectiveness of the proposed framework for HPP operation. Simulation results highlight that the proposed framework not only can help operators schedule HPP similar to Congens in varying weather conditions but can also maintain the frequency reliability of the system.

frequency stability↗

An Integrated Framework for Risk Assessment of High Safety Significant Safety-related Digital Instrumentation and Control Systems in Nuclear Power Plants: Methodology and Demonstration

This report documents the activities performed by Idaho National Laboratory (INL) during Fiscal Year (FY) 2022 for the U.S. Department of Energy (DOE) Light Water Reactor Sustainability (LWRS) Program, Risk Informed Systems Analysis (RISA) Pathway, digital instrumentation and control (DI&C) risk assessment project. In FY 2019, the RISA Pathway initiated a project to develop a risk assessment strategy for delivering a technical basis to support effective and secure DI&C technologies for digital upgrades/designs. A framework was proposed for this strategy, which aims to (1) provide a best-estimate, risk-informed capability to quantitatively and accurately estimate the risk impact of plant modernization, considering the introduction of high safety-significant safety-related (HSSSR) DI&C systems, (2) support and supplement existing risk-informed DI&C design guides by providing quantitative risk information and evidence, (3) offer a capability of design architecture evaluation of various DI&C systems, (4) assure the long-term safety and reliability of HSSSR DI&C systems, and (5) reduce uncertainty in costs and support integration of DI&C systems in the plant. To achieve these technical goals, the framework provides a means to address relevant technical issues by: (1) defining a risk-informed analysis process for DI&C upgrade, that integrates hazard analysis, reliability analysis, and consequence analysis, (2) applying risk-informed tools to address common cause failures (CCFs) and quantify corresponding failure probabilities for DI&C technologies, particularly software CCFs, (3) evaluating the impact of digital failures at the component level, system level, and plant level, and (4) providing insights and suggestions on designs to manage the risks, thus to support the development and deployment of advanced DI&C technologies on nuclear power plant (NPPs).

21 SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLANTS↗

Accounting for Point Estimate Uncertainty in Space Systems Reliability and Risk Analysis

Understanding and accounting for uncertainty in risk analysis is a critical step in the management and communication of risk in engineered systems. The component and system-level analysis to determine the probability of a negative outcome and its consequence is often quantified by a point estimate. Many Program and Enterprise decisions involving technical concerns and issues rely on reliability engineering activities to produce quantified risk analysis to inform the decision making process. At NASA, it is common to use a Probabilistic Risk Analysis (PRA) to inform the overall risk to Loss of Mission or Loss of Crew that involves integration across all spacecraft subsystem fault trees to produce an overall probability of mission failure. The point estimate is an estimate of this overall probability and is an immediate result of a fault tree model. It is the result of a model where the probability of each event is taken to be equal to its mean. The value provides an approximation of the overall mean without running any uncertainty calculations (e.g., no sampling). Using only the point estimate can lead to a false sense of precision and the point estimate may not match the resulting mean when uncertainty is taken into consideration. This paper will explore five conditions that can cause the PRA model mean to diverge from the point estimate and will provide engineers and managers insight into the importance of understanding uncertainty in the elements of PRA models.

Paul J Collier↗

Congestion Management Solutions for Enhanced Distribution System Operations with Aggregated Distribution Grid Resources Providing Grid Services and Market Participation

Microgrids and other aggregations of distribution grid resources (DGRs) are poised to actively participate in electricity markets and provide essential grid services in the coming years. In fact, DGRs already play such a role through behind-the-meter (BTM) demand response programs and small-scale BTM dispatchable generation initiatives. At the same time, the rapid growth of artificial intelligence (AI) and cryptocurrency datacenters imposes significant, often unpredictable, demands on the power distribution system. Aggregated DGRs can serve as flexible resources that help mitigate these pressures by using available transmission and distribution capacity more efficiently, supporting resource adequacy and other reliability services, and providing bridge strategies while long-term transmission infrastructure is being developed. The impacts this activity will have on distribution networks are not fully understood and could present significant challenges for distribution utilities due to capacity constraints and the need for congestion management. Technical issues include reverse power flow, variability and possible degradation of equipment integrity, voltage violations, and customer power quality concerns. These issues will likely intensify as electricity market operators across the United States implement Federal Energy Regulatory Commission Order 2222 over the next few years.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Application of Manufacturing Quality Management Principles to PV System Installations

To help SETO/DOE achieve its goals, the IBTS team proposed a project addressing system reliability by improving installation standards and quality management. The proposed approach was designed to help achieve measurable reductions in installation defect density and improvements in the performance of PV systems by optimizing design and installation of residential and commercial PV systems. This approach addressed the soft costs associated with installations and quality management. The project demonstrated improved system reliability and reduced PV system installation costs. The software developed improved operations, decreased risk, and increased the overall value of PV systems across their lifecycle. The project used several data collection methods, including extensive industry surveys, face-to-face high-level interviews at industry conferences, stakeholder teleconferences, and in-depth interviews conducted by IBTS staff. Results from the research found the industry needs a uniform assessment method for national providers to be more efficient; the software should support both code officials and installers; most industry stakeholders would find value in a centralized software system that allows them to collect, report, and review information on in-process and completed solar installations; and mobile solutions that bridge existing knowledge gaps with inspectors and integrate with existing methodologies (such as permitting software) are of great value. The software solution developed is web-based, allowing for national access, and is built on a Google Firebase platform that can handle significant users and data. It can be used onsite or remotely, allowing for code compliance to continue despite ongoing pandemic related delays or shutdowns for local economies. The information provided by the software tool allows users to uniformly assess a system for compliance and use that aggregated data to identify training topics or create internal process designed to improving issues and reducing occurrence. This solution has multiple benefits in managing quality at time of use and promoting an increase in future safety and quality through education. Perhaps most importantly, this software increases public safety by ensuring compliance of installed systems and allows for local AHJs to remotely engage specialized and qualified solar specific expertise for oversite of the installation in their jurisdictions. Data analysis provides the quality feedback loop identifying the root cause of failure and drives installation practices to improve through training and education, resulting in systems with higher performance, greater reliability, and reduced operations and maintenance costs. With the successful completion of this project, the industry can expect reduced soft costs and increased performance and safety and will ultimately benefit from longer performing systems that cost less to operate.

14 SOLAR ENERGY↗

Integrated Software Health Management for Aircraft GN and C

Modern aircraft rely heavily on dependable operation of many safety-critical software components. Despite careful design, verification and validation (V&V), on-board software can fail with disastrous consequences if it encounters problematic software/hardware interaction or must operate in an unexpected environment. We are using a Bayesian approach to monitor the software and its behavior during operation and provide up-to-date information about the health of the software and its components. The powerful reasoning mechanism provided by our model-based Bayesian approach makes reliable diagnosis of the root causes possible and minimizes the number of false alarms. Compilation of the Bayesian model into compact arithmetic circuits makes SWHM feasible even on platforms with limited CPU power. We show initial results of SWHM on a small simulator of an embedded aircraft software system, where software and sensor faults can be injected.

Schumann, Johann↗

A Simplified Model of VIPER Thermal Management System. Part II: Integrated Vehicle

NASA’s Volatiles Investigating Polar Exploration Rover (VIPER) thermal management system (TMS) relies on four loop heat pipes (LHPs) to transport electronic waste heat to the vehicle cooling radiative surface and avoid overheating. The TMS has also ten constant conductance heat pipes (CCHPs) dedicated to balance the thermal load within the internal environment where the avionics boxes are mounted, also called warm electronic box (WEB), and to transport the heat from two of the science payload instruments. The TMS also uses two thermal straps to thermally link the batteries to the WEB. These thermal components, in addition to heaters, thermostat, multi-layer insulation (MLIs), and isolators forms the core of the VIPER TMS. The complex heat transport balance managed by the TMS is challenging to characterize and model. The more fidelity and granularity of a model, the more costly the computational resources needed and the longer the simulation and modeling time. When the priority is to provide quick but reliable assessments of the thermal performance or real time thermal feedback for training of console operators, simplified modeling tools are needed. To satisfy that need, this paper describes the effort to develop and correlate a model of VIPER TMS based on control volume approach. The correlation effort in particular focuses on hibernation, cold thermal balance, and hot thermal balance data from the integrated vehicle thermal vacuum (TVAC) test. Thus, the correlated model captures the heat leaks during hibernations and the performance at two extremes, bounding, operating scenarios.

Loop Heat Pipe↗