Search NASA⌕ Search

SEARCH · Search NASA

Results for “faults”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 289 records · Page 16

Fault diagnosis for the Space Shuttle main engine

A conceptual design of a model-based fault detection and diagnosis system is developed for the Space Shuttle main engine. The design approach consists of process modeling, residual generation, and fault detection and diagnosis. The engine is modeled using a discrete time, quasilinear state-space representation. Model parameters are determined by identification. Residuals generated from the model are used by a neural network to detect and diagnose engine component faults. Fault diagnosis is accomplished by training the neural network to recognize the pattern of the respective fault signatures. Preliminary results for a failed valve, generated using a full, nonlinear simulation of the engine, are presented. These results indicate that the developed approach can be used for fault detection and diagnosis. The results also show that the developed model is an accurate and reliable predictor of the highly nonlinear and very complex engine.

Duyar, Ahmet↗

Global positioning system measurements of deformations associated with the 1987 Superstition Hills earthquake - Evidence for conjugate faulting

Large station displacements observed from Imperial Valley Global Positioning System (GPS) compaigns are attributed to the November 24, 1987 Superstition Hills earthquake sequence. Thirty sites from a 42 station GPS network established in 1986 were reoccupied during 1988 and/or 1990. Displacements at three sites within 3 kilometers of the surface rupture approach 0.5 m. Eight additional stations within 20 km of the seismic zone are displaced at least 10 cm. This is the first occurrence of a large earthquake (M(sub S) 6.6) within a preexisting GPS network. Best-fitting uniform slip models of rectangular dislocations in an elastic half-space indicate 130 + or - 8 cm right-lateral displacement along the northwest-trending Superstition Hills fault and 30 + or - 10 cm left-lateral displacement along the conjugate northeast-trending Elmore Ranch fault. The geodetic moments are 9.4 x 10 (exp 25) dyne-cm and 2.3 x 10 (exp 25) dyne-cm for the Superstition Hills and Elmore Ranch faults, respectively, consistent with teleseismic source parameters. The data also suggest the post seismic slip along the Superstition Hills fault is concentrated at shallow depths. Distributed slip solutions using Singular Value Decomposition indicate near uniform displacement along the Elmore Ranch fault and concentrated slip to the northwest and southeast along the Superstition Hills fault. A significant component of non-seismic displacement is observed across the Imperial Valley, which is attributed in part to interseismic plate-boundary deformation.

Larsen, Shawn↗

Large transient fault current test of an electrical roll ring

The Space Station Freedom uses precision rotary gimbals to provide for sun tracking of its photoelectric arrays. Electrical power, command signals, and data are transferred across the gimbals by roll rings. Roll rings have been shown to be capable of highly efficient electrical transmission and long life, through tests conducted at the NASA Lewis Research Center and Honeywell's Satellite and Space Systems Division in Phoenix, AZ. Large potential fault currents inherent to the power system's DC distribution architecture have brought about the need to evaluate the effects of large transient fault currents on roll rings. A test recently conducted at Lewis subjected a roll ring to a simulated worst case space station electrical fault. The system model used to obtain the fault profile is described, along with details of the reduced order circuit that was used to simulate the fault. Test results comparing roll ring performance before and after the fault are also presented.

Yenni, Edward J.↗

Fault identification using multidisciplinary techniques at the Mars/Uranus Station antenna sites

A fault investigation was performed at the Mars and Uranus antenna sites at the Goldstone Deep Space Communications Complex in the Mojave desert. The Mars/Uranus Station consists of two large-diameter reflector antennas used for communication and control of deep-space probes and other missions. The investigation included interpretation of Landsat thematic mapper scenes, side-looking airborne radar transparencies, and both color-infrared and black-and-white aerial photography. Four photolineaments suggestive of previously undocumented faults were identified. Three generally discrete morphostratigraphic alluvial-fan deposits were also recognized and dated using geomorphic and soil stratigraphic techniques. Fourteen trenches were excavated across the four lineaments; the trenches show that three of the photolineaments coincide with faults. The last displacement of two of the faults occurred between about 12,000 and 35,000 years ago. The third fault was judged to be older than 12,000 years before present (ybp), although uncertainty remains. None of the surface traces of the three faults crosses under existing antennas or structures; however, their potential activity necessitates appropriate seismic retrofit designs and loss-prevention measures to mitigate potential earthquake damage to facilities and structures.

Santo, D. S.↗

A fault-tolerant intelligent robotic control system

This paper describes the concept, design, and features of a fault-tolerant intelligent robotic control system being developed for space and commercial applications that require high dependability. The comprehensive strategy integrates system level hardware/software fault tolerance with task level handling of uncertainties and unexpected events for robotic control. The underlying architecture for system level fault tolerance is the distributed recovery block which protects against application software, system software, hardware, and network failures. Task level fault tolerance provisions are implemented in a knowledge-based system which utilizes advanced automation techniques such as rule-based and model-based reasoning to monitor, diagnose, and recover from unexpected events. The two level design provides tolerance of two or more faults occurring serially at any level of command, control, sensing, or actuation. The potential benefits of such a fault tolerant robotic control system include: (1) a minimized potential for damage to humans, the work site, and the robot itself; (2) continuous operation with a minimum of uncommanded motion in the presence of failures; and (3) more reliable autonomous operation providing increased efficiency in the execution of robotic tasks and decreased demand on human operators for controlling and monitoring the robotic servicing routines.

Marzwell, Neville I.↗

FOCUS - An experimental environment for fault sensitivity analysis

FOCUS, a simulation environment for conducting fault-sensitivity analysis of chip-level designs, is described. The environment can be used to evaluate alternative design tactics at an early design stage. A range of user specified faults is automatically injected at runtime, and their propagation to the chip I/O pins is measured through the gate and higher levels. A number of techniques for fault-sensitivity analysis are proposed and implemented in the FOCUS environment. These include transient impact assessment on latch, pin and functional errors, external pin error distribution due to in-chip transients, charge-level sensitivity analysis, and error propagation models to depict the dynamic behavior of latch errors. A case study of the impact of transient faults on a microprocessor-based jet-engine controller is used to identify the critical fault propagation paths, the module most sensitive to fault propagation, and the module with the highest potential for causing external errors.

Choi, Gwan S.↗

Fault management for data systems

Issues related to automating the process of fault management (fault diagnosis and response) for data management systems are considered. Substantial benefits are to be gained by successful automation of this process, particularly for large, complex systems. The use of graph-based models to develop a computer assisted fault management system is advocated. The general problem is described and the motivation behind choosing graph-based models over other approaches for developing fault diagnosis computer programs is outlined. Some existing work in the area of graph-based fault diagnosis is reviewed, and a new fault management method which was developed from existing methods is offered. Our method is applied to an automatic telescope system intended as a prototype for future lunar telescope programs. Finally, an application of our method to general data management systems is described.

Boyd, Mark A.↗

Wrinkle ridges, reverse faulting, and the depth penetration of lithospheric stress in lunae planum, Mars

Tectonic features on a planetary surface are commonly used as constraints on models to determine the state of stress at the time the features formed. Quantitative global stress models applied to understand the formation of the Tharsis province on Mars constrained by observed tectonics have calculated stresses at the surface of a thin elastic shell and have neglected the role of vertical structure in influencing the predicted pattern of surface deformation. Wrinkle ridges in the Lunae Planum region of Mars form a conentric pattern of regularly spaced features in the eastern and southeastern part of Tharsis; they are formed due to compressional stresses related to the response of the Martian lithosphere to the Tharsis bulge. As observed in the exposures of valley walls in areas such as the Kasei Valles, the surface plains unit is underlain by an unconsolidated impact-generated megaregolith that grades with depth into structurally competent lithospheric basement. The ridges have alternatively been hypothesized to reflect deformation restricted to the surface plains unit ('thin skinned deformation') and deformation that includes the surface unit, megaregolith and basement lithosphere ('thick skinned deformation'). We have adopted a finite element approach to quantify the nature of deformation associated with the development of wrinkle ridges in a vertically stratified elastic lithosphere. We used the program TECTON, which contains a slippery node capability that allowed us to explicitly take into account the presence of reverse faults believed to be associated with the ridges. In this study we focused on the strain field in the vicinity of a single ridge when slip occurs along the fault. We considered two initial model geometries. In the first, the reverse fault was assumed to be in the surface plains unit, and in the second the initial fault was located in lithospheric basement, immediately beneath the weak megaregolith. We are interested in the conditions underwhich strain in the surface layer and basement either penetrates or fails to penetrate through the megaregolith. We thus address the conditions required for an initial basement fault to propagate through the megaregolith to the surface, as well as the effect of the megareolith on the strain tensor in the vicinity of a fault that nucleates in the surface plains unit.

Zuber, M. T.↗

Detection of feed-through faults in CMOS storage elements

In testing sequential circuits, internal faults in the storage elements (SE's) are sometimes modeled as stuck-at faults in the combinational circuits surrounding the SE. The detection of some transistor-level faults that cannot be modeled as stuck-at are considered. These feed-through faults cause the cell to become either data-feed-through, which makes the cell combinational, or clock-feed-through, which causes the clock signal or its complement to appear at the output. Under such faults, the cell does not function as a memory element. Here it is shown that such faults may or may not be detected depending on delays involved. Conditions under which race-ahead occurs are identified.

Al-Assadi, Waleed K.↗

Application of fault detection techniques to spiral bevel gear fatigue data

Results of applying a variety of gear fault detection techniques to experimental data is presented. A spiral bevel gear fatigue rig was used to initiate a naturally occurring fault and propagate the fault to a near catastrophic condition of the test gear pair. The spiral bevel gear fatigue test lasted a total of eighteen hours. At approximately five and a half hours into the test, the rig was stopped to inspect the gears for damage, at which time a small pit was identified on a tooth of the pinion. The test was then stopped an additional seven times throughout the rest of the test in order to observe and document the growth and propagation of the fault. The test was ended when a major portion of a pinion tooth broke off. A personal computer based diagnostic system was developed to obtain vibration data from the test rig, and to perform the on-line gear condition monitoring. A number of gear fault detection techniques, which use the signal average in both the time and frequency domain, were applied to the experimental data. Among the techniques investigated, two of the recently developed methods appeared to be the first to react to the start of tooth damage. These methods continued to react to the damage as the pitted area grew in size to cover approximately 75% of the face width of the pinion tooth. In addition, information gathered from one of the newer methods was found to be a good accumulative damage indicator. An unexpected result of the test showed that although the speed of the rig was held to within a band of six percent of the nominal speed, and the load within eighteen percent of nominal, the resulting speed and load variations substantially affected the performance of all of the gear fault detection techniques investigated.

Zakrajsek, James J.↗

Software fault tolerance in computer operating systems

This chapter provides data and analysis of the dependability and fault tolerance for three operating systems: the Tandem/GUARDIAN fault-tolerant system, the VAX/VMS distributed system, and the IBM/MVS system. Based on measurements from these systems, basic software error characteristics are investigated. Fault tolerance in operating systems resulting from the use of process pairs and recovery routines is evaluated. Two levels of models are developed to analyze error and recovery processes inside an operating system and interactions among multiple instances of an operating system running in a distributed environment. The measurements show that the use of process pairs in Tandem systems, which was originally intended for tolerating hardware faults, allows the system to tolerate about 70% of defects in system software that result in processor failures. The loose coupling between processors which results in the backup execution (the processor state and the sequence of events occurring) being different from the original execution is a major reason for the measured software fault tolerance. The IBM/MVS system fault tolerance almost doubles when recovery routines are provided, in comparison to the case in which no recovery routines are available. However, even when recovery routines are provided, there is almost a 50% chance of system failure when critical system jobs are involved.

Iyer, Ravishankar K.↗

Spacecraft fault tolerance: The Magellan experience

Interplanetary and earth orbiting missions are now imposing unique fault tolerant requirements upon spacecraft design. Mission success is the prime motivator for building spacecraft with fault tolerant systems. The Magellan spacecraft had many such requirements imposed upon its design. Magellan met these requirements by building redundancy into all the major subsystem components and designing the onboard hardware and software with the capability to detect a fault, isolate it to a component, and issue commands to achieve a back-up configuration. This discussion is limited to fault protection, which is the autonomous capability to respond to a fault. The Magellan fault protection design is discussed, as well as the developmental and flight experiences and a summary of the lessons learned.

Kasuda, Rick↗

Systematic Underestimation of Earthquake Magnitudes from Large Intracontinental Reverse Faults: Historical Ruptures Break Across Segment Boundaries

Because most large-magnitude earthquakes along reverse faults have such irregular and complicated rupture patterns, reverse-fault segments defined on the basis of geometry alone may not be very useful for estimating sizes of future seismic sources. Most modern large ruptures of historical earthquakes generated by intracontinental reverse faults have involved geometrically complex rupture patterns. Ruptures across surficial discontinuities and complexities such as stepovers and cross-faults are common. Specifically, segment boundaries defined on the basis of discontinuities in surficial fault traces, pronounced changes in the geomorphology along strike, or the intersection of active faults commonly have not proven to be major impediments to rupture. Assuming that the seismic rupture will initiate and terminate at adjacent major geometric irregularities will commonly lead to underestimation of magnitudes of future large earthquakes.

Rubin, C. M.↗

Sequential Testing Algorithms for Multiple Fault Diagnosis

In this paper, we consider the problem of constructing optimal and near-optimal test sequencing algorithms for multiple fault diagnosis. The computational complexity of solving the optimal multiple-fault isolation problem is super-exponential, that is, it is much more difficult than the single-fault isolation problem, which, by itself, is NP-hard. By employing concepts from information theory and AND/OR graph search, we present several test sequencing algorithms for the multiple fault isolation problem. These algorithms provide a trade-off between the degree of suboptimality and computational complexity. Furthermore, we present novel diagnostic strategies that generate a diagnostic directed graph (digraph), instead of a diagnostic tree, for multiple fault diagnosis. Using this approach, the storage complexity of the overall diagnostic strategy reduces substantially. The algorithms developed herein have been successfully applied to several real-world systems. Computational results indicate that the size of a multiple fault strategy is strictly related to the structure of the system.

Shakeri, Mojdeh↗

The Design of a Fault-Tolerant COTS-Based Bus Architecture for Space Applications

The high-performance, scalability and miniaturization requirements together with the power, mass and cost constraints mandate the use of commercial-off-the-shelf (COTS) components and standards in the X2000 avionics system architecture for deep-space missions. In this paper, we report our experiences and findings on the design of an IEEE 1394 compliant fault-tolerant COTS-based bus architecture. While the COTS standard IEEE 1394 adequately supports power management, high performance and scalability, its topological criteria impose restrictions on fault tolerance realization. To circumvent the difficulties, we derive a "stack-tree" topology that not only complies with the IEEE 1394 standard but also facilitates fault tolerance realization in a spaceborne system with limited dedicated resource redundancies. Moreover, by exploiting pertinent standard features of the 1394 interface which are not purposely designed for fault tolerance, we devise a comprehensive set of fault detection mechanisms to support the fault-tolerant bus architecture.

Chau, Savio N.↗

Software Fault Tolerance: A Tutorial

Because of our present inability to produce error-free software, software fault tolerance is and will continue to be an important consideration in software systems. The root cause of software design errors is the complexity of the systems. Compounding the problems in building correct software is the difficulty in assessing the correctness of software for highly complex systems. After a brief overview of the software development processes, we note how hard-to-detect design faults are likely to be introduced during development and how software faults tend to be state-dependent and activated by particular input sequences. Although component reliability is an important quality measure for system level analysis, software reliability is hard to characterize and the use of post-verification reliability estimates remains a controversial issue. For some applications software safety is more important than reliability, and fault tolerance techniques used in those applications are aimed at preventing catastrophes. Single version software fault tolerance techniques discussed include system structuring and closure, atomic actions, inline fault detection, exception handling, and others. Multiversion techniques are based on the assumption that software built differently should fail differently and thus, if one of the redundant versions fails, it is expected that at least one of the other versions will provide an acceptable output. Recovery blocks, N-version programming, and other multiversion techniques are reviewed.

Torres-Pomales, Wilfredo↗

Aircraft Engine Sensor/Actuator/Component Fault Diagnosis Using a Bank of Kalman Filters

In this report, a fault detection and isolation (FDI) system which utilizes a bank of Kalman filters is developed for aircraft engine sensor and actuator FDI in conjunction with the detection of component faults. This FDI approach uses multiple Kalman filters, each of which is designed based on a specific hypothesis for detecting a specific sensor or actuator fault. In the event that a fault does occur, all filters except the one using the correct hypothesis will produce large estimation errors, from which a specific fault is isolated. In the meantime, a set of parameters that indicate engine component performance is estimated for the detection of abrupt degradation. The performance of the FDI system is evaluated against a nonlinear engine simulation for various engine faults at cruise operating conditions. In order to mimic the real engine environment, the nonlinear simulation is executed not only at the nominal, or healthy, condition but also at aged conditions. When the FDI system designed at the healthy condition is applied to an aged engine, the effectiveness of the FDI system is impacted by the mismatch in the engine health condition. Depending on its severity, this mismatch can cause the FDI system to generate incorrect diagnostic results, such as false alarms and missed detections. To partially recover the nominal performance, two approaches, which incorporate information regarding the engine s aging condition in the FDI system, will be discussed and evaluated. The results indicate that the proposed FDI system is promising for reliable diagnostics of aircraft engines.

Kobayashi, Takahisa↗

Automated fault-management in a simulated spaceflight micro-world

BACKGROUND: As human spaceflight missions extend in duration and distance from Earth, a self-sufficient crew will bear far greater onboard responsibility and authority for mission success. This will increase the need for automated fault management (FM). Human factors issues in the use of such systems include maintenance of cognitive skill, situational awareness (SA), trust in automation, and workload. This study examine the human performance consequences of operator use of intelligent FM support in interaction with an autonomous, space-related, atmospheric control system. METHODS: An expert system representing a model-base reasoning agent supported operators at a low level of automation (LOA) by a computerized fault finding guide, at a medium LOA by an automated diagnosis and recovery advisory, and at a high LOA by automate diagnosis and recovery implementation, subject to operator approval or veto. Ten percent of the experimental trials involved complete failure of FM support. RESULTS: Benefits of automation were reflected in more accurate diagnoses, shorter fault identification time, and reduced subjective operator workload. Unexpectedly, fault identification times deteriorated more at the medium than at the high LOA during automation failure. Analyses of information sampling behavior showed that offloading operators from recovery implementation during reliable automation enabled operators at high LOA to engage in fault assessment activities CONCLUSIONS: The potential threat to SA imposed by high-level automation, in which decision advisories are automatically generated, need not inevitably be counteracted by choosing a lower LOA. Instead, freeing operator cognitive resources by automatic implementation of recover plans at a higher LOA can promote better fault comprehension, so long as the automation interface is designed to support efficient information sampling.

NASA Discipline Space Human Factors↗