Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “hardware failure”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

Knowledge representation and user interface concepts to support mixed-initiative diagnosis

The Remote Maintenance Monitoring System (RMMS) provides automated support for the maintenance and repair of ModComp computer systems used in the Launch Processing System (LPS) at Kennedy Space Center. RMMS supports manual and automated diagnosis of intermittent hardware failures, providing an efficient means for accessing and analyzing the data generated by catastrophic failure recovery procedures. This paper describes the design and functionality of the user interface for interactive analysis of memory dump data, relating it to the underlying declarative representation of memory dumps.

Sobelman, Beverly H.↗

REDEX - The ranging equipment diagnostic expert system

REDEX, an advanced prototype expert system that diagnoses hardware failures in the Ranging Equipment (RE) at NASA's Ground Network tracking stations is described. REDEX will help the RE technician identify faulty circuit cards or modules that must be replaced, and thereby reduce troubleshooting time. It features a highly graphical user interface that uses color block diagrams and layout diagrams to illustrate the location of a fault. A semantic network knowledge representation technique was used to model the design structure of the RE. A catalog of generic troubleshooting rules was compiled to represent heuristics that are applied in diagnosing electronic equipment. Specific troubleshooting rules were identified to represent additional diagnostic knowledge that is unique to the RE. Over 50 generic and 250 specific troubleshooting rules have been derived. REDEX is implemented in Prolog on an IBM PC AT-compatible workstation. Block diagram graphics displays are color-coded to identify signals that have been monitored or inferred to have nominal values, signals that are out of tolerance, and circuit cards and functions that are diagnosed as faulty. A hypertext-like scheme is used to allow the user to easily navigate through the space of diagrams and tables. Over 50 graphic and tabular displays have been implemented. REDEX is currently being evaluated in a stand-alone mode using simulated RE fault scenarios. It will soon be interfaced to the RE and tested in an online environment. When completed and fielded, REDEX will be a concrete example of the application of expert systems technology to the problem of improving performance and reducing the lifecycle costs of operating NASA's communications networks in the 1990s.

Luczak, Edward C.↗

Fault-tolerant multichannel demultiplexer subsystems

Fault tolerance in future processing and switching communication satellites is addressed by showing new methods for detecting hardware failures in the first major subsystem, the multichannel demultiplexer. An efficient method for demultiplexing frequency slotted channels uses multirate filter banks which contain fast Fourier transform processing. All numerical processing is performed at a lower rate commensurate with the small bandwidth of each bandbase channel. The integrity of the demultiplexing operations is protected by using real number convolutional codes to compute comparable parity values which detect errors at the data sample level. High rate, systematic convolutional codes produce parity values at a much reduced rate, and protection is achieved by generating parity values in two ways and comparing them. Parity values corresponding to each output channel are generated in parallel by a subsystem, operating even slower and in parallel with the demultiplexer that is virtually identical to the original structure. These parity calculations may be time shared with the same processing resources because they are so similar.

Redinbo, Robert↗

Evaluation of spacecraft product assurance requirements and flight performance history

The Jet Propulsion Laboratory (JPL) has undertaken a task to relate long duration, space flight hardware performance to specific product assurance requirements established during the hardware development process. This paper describes the approach that JPL is using to correlate in-flight and ground test hardware failures to critical product assurance practices implemented on a given flight project. The first step in this effort has been to collect, and convert into a convenient format, in-flight problem, failure, and anomaly data. A characterization of a subset of this anomaly data base, focusing on anomaly causes, time dependence, and subsystem affected is presented.

Gonzalez, Charles C.↗

FORTH as the basis for an integrated operations environment for a Space Shuttle scientific experiment

Over a period of three years, a FORTH-based system was developed by JPL for the operations of a major scientific instrument onboard the Space Shuttle. The software had to meet a very demanding operations environment where the interactiveness of the software was not merely desirable but essential to the success of the mission. Forth was chosen for its capability of integrating divergent software needs into an interactive package. The mission flown in October 1984 was beset with numerous hardware failures and challenged the capability of the system to its fullest.

Harris, Henry M.↗

Progressive retry for software error recovery in distributed systems

In this paper, we describe a method of execution retry for bypassing software errors based on checkpointing, rollback, message reordering and replaying. We demonstrate how rollback techniques, previously developed for transient hardware failure recovery, can also be used to recover from software faults by exploiting message reordering to bypass software errors. Our approach intentionally increases the degree of nondeterminism and the scope of rollback when a previous retry fails. Examples from our experience with telecommunications software systems illustrate the benefits of the scheme.

Wang, Yi-Min↗

High-performance reactionless scan mechanism

A high-performance reactionless scan mirror mechanism was developed for space applications to provide thermal images of the Earth. The design incorporates a unique mechanical means of providing reactionless operation that also minimizes weight, mechanical resonance operation to minimize power, combined use of a single optical encoder to sense coarse and fine angular position, and a new kinematic mount of the mirror. A flex pivot hardware failure and current project status are discussed.

Williams, Ellen I.↗

A Scheduling Algorithm for Replicated Real-Time Tasks

We present an algorithm for scheduling real-time periodic tasks on a multiprocessor system under fault-tolerant requirement. Our approach incorporates both the redundancy and masking technique and the imprecise computation model. Since the tasks in hard real-time systems have stringent timing constraints, the redundancy and masking technique are more appropriate than the rollback techniques which usually require extra time for error recovery. The imprecise computation model provides flexible functionality by trading off the quality of the result produced by a task with the amount of processing time required to produce it. It therefore permits the performance of a real-time system to degrade gracefully. We evaluate the algorithm by stochastic analysis and Monte Carlo simulations. The results show that the algorithm is resilient under hardware failures.

Yu, Albert C.↗

Logic Design Pathology and Space Flight Electronics

This paper presents a look at logic design from early in the US Space Program and examines faults in recent logic designs. Most examples are based on flight hardware failures and analysis of new tools and techniques. The paper is presented in viewgraph form.

Katz, Richard B.↗

ESTAR Measurements During SGP-99

The synthetic aperture radiometer, ESTAR, provided L-band brightness temperature maps of the experiment site during the Southern Great Plains Experiment in 1999. ESTAR flew on the NASA P-3 aircraft at an altitude of 7.6 km and mapped a swath about 50 km wide and about 300 km long. The area mapped extended west from Oklahoma City to El Reno and north from the Little Washita River watershed to the Kansas border. The flight lines and mapping were done in a manner similar to that used in 1997 during SGP-97. ESTAR is a prototype designed to develop the technology of aperture synthesis for passive microwave remote sensing. It is actually a hybrid that uses real aperture to obtain resolution along track and synthetic aperture to obtain resolution across track. ESTAR operates in the band at 1.413 GHz set aside for passive use and images at horizontal polarization in the equivalent of a cross-track scan. The scan is done in software as part of image reconstruction. Calibration consists of blackbody observed before and after each flight and a local body of water (Lake Kaw). ESTAR arrived in Oklahoma on July 7, 1999 and flew mapping missions on July 8,9, 11, 14, 15, 19 and 20. The aircraft was down on July 10 by design and was down again on July 12-13 and 16-18 due to mechanical problems. ESTAR data for July I I is questionable because of a hardware failure. The brightness temperature maps reflect the patterns of soil moisture observed in the SGP99 study area and compare well with measurements made by ESTAR in other experiments in this region (e.g. SGP -97 and Washita-92).

LeVine, D. M.↗

Flight Test of Propulsion Monitoring and Diagnostic System

The objective of this program was to perform flight tests of the propulsion monitoring and diagnostic system (PMDS) technology concept developed by Honeywell under the NASA Advanced General Aviation Transport Experiment (AGATE) program. The PMDS concept is intended to independently monitor the performance of the engine, providing continuous status to the pilot along with warnings if necessary as well as making the data available to ground maintenance personnel via a special interface. These flight tests were intended to demonstrate the ability of the PMDS concept to detect a class of selected sensor hardware failures, and the ability to successfully model the engine for the purpose of engine diagnosis.

Gabel, Steve↗

Development and Evaluation of Fault-Tolerant Flight Control Systems

The research is concerned with developing a new approach to enhancing fault tolerance of flight control systems. The original motivation for fault-tolerant control comes from the need for safe operation of control elements (e.g. actuators) in the event of hardware failures in high reliability systems. One such example is modem space vehicle subjected to actuator/sensor impairments. A major task in flight control is to revise the control policy to balance impairment detectability and to achieve sufficient robustness. This involves careful selection of types and parameters of the controllers and the impairment detecting filters used. It also involves a decision, upon the identification of some failures, on whether and how a control reconfiguration should take place in order to maintain a certain system performance level. In this project new flight dynamic model under uncertain flight conditions is considered, in which the effects of both ramp and jump faults are reflected. Stabilization algorithms based on neural network and adaptive method are derived. The control algorithms are shown to be effective in dealing with uncertain dynamics due to external disturbances and unpredictable faults. The overall strategy is easy to set up and the computation involved is much less as compared with other strategies. Computer simulation software is developed. A serious of simulation studies have been conducted with varying flight conditions.

Song, Yong D.↗

The Wallops Flight Facility Rapid Response Range Operations Initiative

While the dominant focus on short response missions has appropriately centered on the launch vehicle and spacecraft, often overlooked or afterthought phases of these missions have been launch site operations and the activities of launch range organizations. Throughout the history of organized spaceflight, launch ranges have been the bane of flight programs as the source of expense, schedule delays, and seemingly endless requirements. Launch Ranges provide three basic functions: (1) provide an appropriate geographical location to meet orbital other mission trajectory requirements, (2) provide project services such as processing facilities, launch complexes, tracking and data services, and expendable products, and (3) assure safety and property protection to participating personnel and third-parties. The challenge with which launch site authorities continuously struggle, is the inherent conflict arising from projects whose singular concern is execution of their mission, and the range s need to support numerous simultaneous customers. So, while tasks carried out by a launch range committed to a single mission pale in comparison to efforts of a launch vehicle or spacecraft provider and could normally be carried out in a matter of weeks, major launch sites have dozens of active projects separate sponsoring organizations. Accommodating the numerous tasks associated with each mission, when hardware failures, weather, maintenance requirements, and other factors constantly conspire against the range resource schedulers, make the launch range as significant an impediment to responsive missions as launch vehicles and their cargo. The obvious solution to the launch site challenge was implemented years ago when the Department of Defense simply established dedicated infrastructure and personnel to dedicated missions, namely an Inter Continental Ballistic Missile. This however proves to be prohibitively expensive for all but the most urgent of applications. So the challenge becomes how can a launch site provide acceptably responsive mission services to a particular customer without dedicating extensive resources and while continuing to serve other projects? NASA's Wallops Flight Facility (WFF) is pursuing solutions to exactly this challenge. NASA, in partnership with the Virginia Commercial Space Flight Authority, has initiated the Rapid Response Range Operations Initiative (R3Ops). R3Ops is a multi-phased effort to incrementally establish and demonstrate increasingly responsive launch operations, with an ultimate goal of providing ELV-class services in a maximum of 7-10 days from initial notification routinely, and shorter schedules possible with committed resources. This target will be pursued within the reality of simultaneous concurrent programs, and ideally, largely independent of specialized flight system configurations. WFF has recently completed Phase 1 of R3Ops, an in-depth collection (through extensive expert interviews) and software modeling of individual steps by various range disciplines. This modeling is now being used to identify existing inefficiencies in current procedures, to identify bottlenecks, and show interdependencies. Existing practices are being tracked to provide a baseline to benchmark against as new procedures are implemented. This paper will describe in detail the philosophies behind WFF's R3Ops, the data collected and modeled in Phase 1, and strategies for meeting responsive launch requirements in a multi-user range environment planned for subsequent phases of this initiative.

Underwood, Bruce E.↗

Standardization Efforts for Mechanical Testing and Design of Advanced Ceramic Materials and Components

Advanced aerospace systems occasionally require the use of very brittle materials such as sapphire and ultra-high temperature ceramics. Although great progress has been made in the development of methods and standards for machining, testing and design of component from these materials, additional development and dissemination of standard practices is needed. ASTM Committee C28 on Advanced Ceramics and ISO TC 206 have taken a lead role in the standardization of testing for ceramics, and recent efforts and needs in standards development by Committee C28 on Advanced Ceramics will be summarized. In some cases, the engineers, etc. involved are unaware of the latest developments, and traditional approaches applicable to other material systems are applied. Two examples of flight hardware failures that might have been prevented via education and standardization will be presented.

Salem, Jonathan A.↗

Hubble Space Telescope Magnetometer and Two-Gyro Control Law Design, Implementation, and On-Orbit Performance

For f i h years, the science mission of the Hubble Space Telescope (HST) required using at least three of six rate gyros for attitude control. In the past, HST has mitigated gyro hardware failures by replacement of the failed units through Space Shuttle Servicing Missions. Following the tragic loss of Space Shuttle Columbia on STS-107, the desire to have a safe haven for astronauts during missions has resulted in the cancellation of all planned maxu14 missions to HST. While a robotic servicing mission is being currently being planned, controlling with alternate sensors to replace failed gyros can extend the HST Science mission until the robotic mission can be performed and extend science at HST s end of life. A two-gym control law has been designed and implemented using magnetometers (Magnetic Sensing System - MSS), fixed head star trackers (FHSTs), and Fine Guidance Sensors (FGSs) to control vehicle rate about the missing gyro axis. The three aforementioned sensors are used in succession to reduce HST boresight jitter to less than 7 milli-arcseconds rms and attitude error to less than 10 milli-arcseconds prior to science imaging. The MSS and 2-Gyro (M2G) control law is used for large angle maneuvers and attitude control during earth occultation of FHSTs and FGSs. The Tracker and 2-Gyro (T2G) control law dampens M2G rates and corrects the majority of attitude error in preparation for guide star acquisition with the FGSs. The Fine Guidance Sensor and 2-Gyro (F2G) control law d a m p T2G rates and controls HST attitude during science imaging. This paper describes the M2G control law. Details of M2G algorithms are presented, including computation of the HST 3-axis attitude error estimate, design of the M2G control law, SISO hear stability analyses, and restrictions on operations to maintain the h d t h and safety requirement of a 10degree maximum attitude error. Results of simulations performed in HSTSIM, a high-fidelity non-linear time domain simulation, are presented to predict HST on-orbit performance in attitude hold and maneuver modes. Simulation results are compared to on-orbit data from M2G flight tests performed in November and December 2004 and February 2005. Flight telemetry, using a currently available third gyro, shows that HST attitude error with the new M2G control law is maintained below the 10-degree requirement, and attitude errors are under 2 degrees for 95% of the time.

Wirzburger, John H.↗

Software Architecture of Sensor Data Distribution In Planetary Exploration

Data from mobile and stationary sensors will be vital in planetary surface exploration. The distribution and collection of sensor data in an ad-hoc wireless network presents a challenge. Irregular terrain, mobile nodes, new associations with access points and repeaters with stronger signals as the network reconfigures to adapt to new conditions, signal fade and hardware failures can cause: a) Data errors; b) Out of sequence packets; c) Duplicate packets; and d) Drop out periods (when node is not connected). To mitigate the effects of these impairments, a robust and reliable software architecture must be implemented. This architecture must also be tolerant of communications outages. This paper describes such a robust and reliable software infrastructure that meets the challenges of a distributed ad hoc network in a difficult environment and presents the results of actual field experiments testing the principles and actual code developed.

Lee, Charles↗

Community Coordinated Modeling Center Support of Operations: Real-Time Simulations and V & V.

In support of Operations Community Coordinated Modeling Center (CCMC) performing validation and verification of space weather models. To identify suitable metrics the CCMC focus on parameters most useful to operations that CCMC resident models can provide. The real time simulations carried out at CCMC are an essential tool to test model performance and stability by using input conditions that may occur in nature at any time. Since 2001, the magnetospheric MHD model BATSRUS has been run in real time using ACE real time data. CCMC staff developed an experimental real-time system that controls uploading of the real-time ACE data, monitors continuous model execution, initiates automatic recovery procedure in case of data gaps or hardware failures, synchronizes BATSRUS and FRC runs, and periodically runs IDL based visualization software.

Kuznetsova, M.↗

Development of a Modified Vacuum Cleaner for Lunar Surface Systems

The National Aeronautics and Space Administration (NASA) mission to expand space exploration will return humans to the Moon with the goal of maintaining a long-term presence. One challenge that NASA will face returning to the Moon is managing the lunar regolith found on the Moon's surface, which will collect on extravehicular activity (EVA) suits and other equipment. Based on the Apollo experience, the issues astronauts encountered with lunar regolith included eye/lung irritation, and various hardware failures (seals, screw threads, electrical connectors and fabric contamination), which were all related to inadequate lunar regolith mitigation. A vacuum cleaner capable of detaching, transferring, and efficiently capturing lunar regolith has been proposed as a method to mitigate the lunar regolith problem in the habitable environment on lunar surface. In order to develop this vacuum, a modified "off-the-shelf" vacuum cleaner has been used to determine detachment efficiency, vacuum requirements, and optimal cleaning techniques to ensure efficient dust removal in habitable lunar surfaces, EVA spacesuits, and air exchange volume. During the initial development of the Lunar Surface System vacuum cleaner, systematic testing was performed with varying flow rates on multiple surfaces (fabrics and metallics), atmospheric (14.7 psia) and reduced pressures (10.2 and 8.3 psia), different vacuum tool attachments, and several vacuum cleaning techniques to determine the performance requirements for the vacuum cleaner. The data recorded during testing was evaluated by calculating percent removal, relative to the retained simulant on the tested surface. In addition, Scanning Electron Microscopy (SEM) imaging was used to determine particle size distribution retained on the surface. The scope of this paper is to explain the initial phase of vacuum cleaner development, including historical Apollo mission data, current state-of-the-art vacuum cleaner technology, and vacuum cleaner testing that has focused on detachment capabilities varying pressure environments.

Toon, Katherine P.↗