Search NASASearch

SEARCH · Search NASA

Results for “multiple faults”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Diagnosing faults in autonomous robot plan execution

A major requirement for an autonomous robot is the capability to diagnose faults during plan execution in an uncertain environment. Many diagnostic researches concentrate only on hardware failures within an autonomous robot. Taking a different approach, the implementation of a Telerobot Diagnostic System that addresses, in addition to the hardware failures, failures caused by unexpected event changes in the environment or failures due to plan errors, is described. One feature of the system is the utilization of task-plan knowledge and context information to deduce fault symptoms. This forward deduction provides valuable information on past activities and the current expectations of a robotic event, both of which can guide the plan-execution inference process. The inference process adopts a model-based technique to recreate the plan-execution process and to confirm fault-source hypotheses. This technique allows the system to diagnose multiple faults due to either unexpected plan failures or hardware errors. This research initiates a major effort to investigate relationships between hardware faults and plan errors, relationships which were not addressed in the past. The results of this research will provide a clear understanding of how to generate a better task planner for an autonomous robot and how to recover the robot from faults in a critical environment.

Lam, Raymond K.

Multiple strike slip faults sets: A case study from the Dead Sea transform

In many strike slip tectonic settings, large rotations of crust blocks about vertical axes have been inferred from paleomagnetic data. These blocks are bounded by sets of parallel faults which presumably accommodate the relative motion between the blocks as regional deformation progress. A mechanical model by Nur et al., (1986) suggests that rotations greater than phi sub c equals 25 to 45 degrees must be accommodated by more than one set of faults, with angle phi sub c between their direction; consequently the sum of the angles between sets must be roughly equal to the total tectonic material rotation. To test this model, the authors investigated the fault geometry and field relation of fault sets in the Mt. Hermon area in northern Israel, where paleomagnetic declination implies data 69 degrees plus or minus 13 degrees counter-clockwise block rotation. The statistical and field relation analysis of over 315 faults shows that the faulting is predominantly right lateral strike slip consisting of three distinct sets. The oldest set strikes 253 degrees, the second oldest set strikes 293 degrees and the youngest strikes 339 degrees. This last direction is consistent also with the current north-south direction of the maximum principle stress axis. The angle phi sub c between the first and second sets is 39 degrees and between the second and third sets 46 degrees, in good agreement with the phi sub c angle predicted from mechanical considerations. The sum of the two angles is 85 degrees, in good agreement with the 69 degrees plus or minus 13 degrees CCW paleomagnetically derived rotation. The results suggest specifically that the sequential development of multiple intersecting fault sets is responsible for the faulting in the Mt. Hermon area; and generally that the model of block rotation with multiple faults provides very good simple rules for analyzing very complex fault patterns.

Ron, Hagai

Integrated analysis of error detection and recovery

An integrated modeling and analysis of error detection and recovery is presented. When fault latency and/or error latency exist, the system may suffer from multiple faults or error propagations which seriously deteriorate the fault-tolerant capability. Several detection models that enable analysis of the effect of detection mechanisms on the subsequent error handling operations and the overall system reliability were developed. Following detection of the faulty unit and reconfiguration of the system, the contaminated processes or tasks have to be recovered. The strategies of error recovery employed depend on the detection mechanisms and the available redundancy. Several recovery methods including the rollback recovery are considered. The recovery overhead is evaluated as an index of the capabilities of the detection and reconfiguration mechanisms.

Shin, K. G.

Software for Fault-Tolerant Matrix Multiplication

Formal Linear Algebra Recovery Environment is a computer program for high-performance, fault-tolerant matrix multiplication. The program is based on an extension of the prior theory and practice of fault-tolerant matrix matrix multiplication of the form C = AB. This extension provides low-overhead methods for detecting errors, not only in C, but also in A and/or B. These methods enable the detection of all errors as long as, in a given case, only one entry in A, B, or C is corrupted. The program also provides for following a low-overhead rollback approach to correct errors once detected. Results of computational experiments have demonstrated that the methods implemented in this program work well in practice while imposing an acceptably low level of overhead, relative to high-performance matrix-multiplication methods that do not afford fault tolerance.

Katz, Daniel

The software-implemented fault tolerance /SIFT/ approach to fault tolerant computing

SIFT is an experimental computer designed for highly reliable flight-control service in advanced air transports. Its development was intended to integrate and demonstrate the latest techniques in fault-tolerant computing. During its development, several new problems of some generality were uncovered and solved. The technology developed for the validation of its design is seen as being perhaps as important as the design itself. The SIFT design is described, as is the way in which the design and its validation were shaped by the requirements of its intended application. Attention is also given to reliability and fault tolerance. The most significant feature of the hardware design is the absence of elements that can generate multiple faults, such as shared clocks or data buses. It is noted that the software is realized in only 800 lines of code, of which 80% are in a high-level language.

Goldberg, J.

Stress field rotation or block rotation: An example from the Lake Mead fault system

The Coulomb criterion, as applied by Anderson (1951), has been widely used as the basis for inferring paleostresses from in situ fault slip data, assuming that faults are optimally oriented relative to the tectonic stress direction. Consequently if stress direction is fixed during deformation so must be the faults. Freund (1974) has shown that faults, when arranged in sets, must generally rotate as they slip. Nur et al., (1986) showed how sufficiently large rotations require the development of new sets of faults which are more favorably oriented to the principal direction of stress. This leads to the appearance of multiple fault sets in which older faults are offset by younger ones, both having the same sense of slip. Consequently correct paleostress analysis must include the possible effect of fault and material rotation, in addition to stress field rotation. The combined effects of stress field rotation and material rotation were investigated in the Lake Meade Fault System (LMFS) especially in the Hoover Dam area. Fault inversion results imply an apparent 60 degrees clockwise (CW) rotation of the stress field since mid-Miocene time. In contrast structural data from the rest of the Great Basin suggest only a 30 degrees CW stress field rotation. By incorporating paleomagnetic and seismic evidence, the 30 degrees discrepancy can be neatly resolved. Based on paleomagnetic declination anomalies, it is inferred that slip on NW trending right lateral faults caused a local 30 degrees counter-clockwise (CCW) rotation of blocks and faults in the Lake Mead area. Consequently the inferred 60 degrees CW rotation of the stress field in the LMFS consists of an actual 30 degrees CW rotation of the stress field (as for the entire Great Basin) plus a local 30 degrees CCW material rotation of the LMFS fault blocks.

Ron, Hagai

Tectonic motion in the western United States inferred from very long baseline interferometry measurements, 1980-1986

Over six years of mobile very long baseline interferometry (VLBI) baseline measurements between 12 sites in the western U.S. were used to infer their velocities relative to the North American plate. These velocities were found to be generally consistent with those determined from geologic data and contemporaneous satellite laser ranging measurements in the same region. The discrepancy between the largest velocities determined from the VLBI measurements of 40-48 mm/yr and the relative plate velocity of 50-56 mm/yr predicted from plate motion models is found to be consistent with a broadened distribution of interseismic strain from cyclic activity on the San Andreas and subsidiary faults. The VLBI data are best explained by a cumulative rate of strike-slip motion near the plate boundary of approximately 48 mm/yr, although exclusion of competing values of 56 and 41 mm/yr is based upon very few data. The rates of offshore fault slip inferred from this study range from about 15 mm/yr in central California to negligible amounts in the San Francisco region. Finite element calculations of multiple fault strain distributions show good agreement with systematic variations in the distribution of shear strain along the San Andreas system, as revealed by previous geodetic measurements.

Kroger, Peter M.

Computing Fault Displacements from Surface Deformations

Simplex is a computer program that calculates locations and displacements of subterranean faults from data on Earth-surface deformations. The calculation involves inversion of a forward model (given a point source representing a fault, a forward model calculates the surface deformations) for displacements, and strains caused by a fault located in isotropic, elastic half-space. The inversion involves the use of nonlinear, multiparameter estimation techniques. The input surface-deformation data can be in multiple formats, with absolute or differential positioning. The input data can be derived from multiple sources, including interferometric synthetic-aperture radar, the Global Positioning System, and strain meters. Parameters can be constrained or free. Estimates can be calculated for single or multiple faults. Estimates of parameters are accompanied by reports of their covariances and uncertainties. Simplex has been tested extensively against forward models and against other means of inverting geodetic data and seismic observations. This work

Lyzenga, Gregory

Real-Time Distributed Embedded Oscillator Operating Frequency Monitoring

A document discusses the utilization of embedded clocks inside of operating network data links as an auxiliary clock source to satisfy local oscillator monitoring requirements. Modem network interfaces, typically serial network links, often contain embedded clocking information of very tight precision to recover data from the link. This embedded clocking data can be utilized by the receiving device to monitor the local oscillator for tolerance to required specifications, often important in high-integrity fault-tolerant applications. A device can utilize a received embedded clock to determine if the local or the remote device is out of tolerance by using a single link. The local device can determine if it is failing, assuming a single fault model, with two or more active links. Network fabric components, containing many operational links, can potentially determine faulty remote or local devices in the presence of multiple faults. Two methods of implementation are described. In one method, a recovered clock can be directly used to monitor the local clock as a direct replacement of an external local oscillator. This scheme is consistent with a general clock monitoring function whereby clock sources are clocking two counters and compared over a fixed interval of time. In another method, overflow/underflow conditions can be used to detect clock relationships for monitoring. These network interfaces often provide clock compensation circuitry to allow data to be transferred from the received (network) clock domain to the internal clock domain. This circuit could be modified to detect overflow/underflow conditions of the buffering required and report a fast or slow receive clock, respectively.

Pollock, Julie

Incipient fault detection study for advanced spacecraft systems

A feasibility study to investigate the application of vibration monitoring to the rotating machinery of planned NASA advanced spacecraft components is described. Factors investigated include: (1) special problems associated with small, high RPM machines; (2) application across multiple component types; (3) microgravity; (4) multiple fault types; (5) eight different analysis techniques including signature analysis, high frequency demodulation, cepstrum, clustering, amplitude analysis, and pattern recognition are compared; and (6) small sample statistical analysis is used to compare performance by computation of probability of detection and false alarm for an ensemble of repeated baseline and faulted tests. Both detection and classification performance are quantified. Vibration monitoring is shown to be an effective means of detecting the most important problem types for small, high RPM fans and pumps typical of those planned for the advanced spacecraft. A preliminary monitoring system design and implementation plan is presented.

Milner, G. Martin

The Soil Moisture Acttive Passive Mission: Fault Protection Performance and Lessons Learned

Fault protection as a discipline involves a collection of flight software logic and operational processes for detecting unacceptable anomalous behavior, responding prior to reaching criticality, restricting the propagation of a failure beyond a fault containment region, and recovering the vehicle back to full or degraded functionality if possible. The System Fault Protection (SFP) design for the SMAP Earth orbiter was put to the test during its 90-day vehicle commissioning activities. During this time, the SFP software autonomously protected the vehicle from multiple faults to critical hardware, and the operations team successfully returned the observatory to its science state. The SFP also performed well in the presence of anomalous behavior below true safety limits by not taking unnecessary response actions, instead allowing the operations team time to monitor the behavior. Certain aspects of the SFP design were modified during operations via both parameter updates and a full flight software update in order to better match the vehicle behavior in the flight environment. An evaluation of the SMAP SFP performance during vehicle Commissioning will be provided in this paper, as well as a set of lessons learned largely focused on visibility, SFP mutability in operations, responses to peripheral device faults, and Safe Mode recovery and design. By capturing some of the knowledge gained during SMAP Commissioning, it is intended that this paper provide guidance for making future System Fault Protection designs more robust and supportive of operations.

Clark, Jessica

Model-Based Diagnostics for Propellant Loading Systems

The loading of spacecraft propellants is a complex, risky operation. Therefore, diagnostic solutions are necessary to quickly identify when a fault occurs, so that recovery actions can be taken or an abort procedure can be initiated. Model-based diagnosis solutions, established using an in-depth analysis and understanding of the underlying physical processes, offer the advanced capability to quickly detect and isolate faults, identify their severity, and predict their effects on system performance. We develop a physics-based model of a cryogenic propellant loading system, which describes the complex dynamics of liquid hydrogen filling from a storage tank to an external vehicle tank, as well as the influence of different faults on this process. The model takes into account the main physical processes such as highly nonequilibrium condensation and evaporation of the hydrogen vapor, pressurization, and also the dynamics of liquid hydrogen and vapor flows inside the system in the presence of helium gas. Since the model incorporates multiple faults in the system, it provides a suitable framework for model-based diagnostics and prognostics algorithms. Using this model, we analyze the effects of faults on the system, derive symbolic fault signatures for the purposes of fault isolation, and perform fault identification using a particle filter approach. We demonstrate the detection, isolation, and identification of a number of faults using simulation-based experiments.

Daigle, Matthew John

Airborne Advanced Reconfigurable Computer System (ARCS)

A digital computer subsystem fault-tolerant concept was defined, and the potential benefits and costs of such a subsystem were assessed when used as the central element of a new transport's flight control system. The derived advanced reconfigurable computer system (ARCS) is a triple-redundant computer subsystem that automatically reconfigures, under multiple fault conditions, from triplex to duplex to simplex operation, with redundancy recovery if the fault condition is transient. The study included criteria development covering factors at the aircraft's operation level that would influence the design of a fault-tolerant system for commercial airline use. A new reliability analysis tool was developed for evaluating redundant, fault-tolerant system availability and survivability; and a stringent digital system software design methodology was used to achieve design/implementation visibility.

Bjurman, B. E.

Error detection process - Model, design, and its impact on computer performance

An analytical model is developed for computer error detection processes and applied to estimate their influence on system performance. Faults in the hardware, not in the design, are assumed to be the potential cause of transition to erroneous states during normal operations. The classification properties and associated recovery methods of error detection are discussed. The probability of obtaining an unreliable result is evaluated, along with the resulting computational loss. Error detection during design is considered and a feasible design space is outlined. Extension of the methods to account for the effects of extant multiple faults is indicated.

Shin, K. G.

Automatically generated acceptance test: A software reliability experiment

This study presents results of a software reliability experiment investigating the feasibility of a new error detection method. The method can be used as an acceptance test and is solely based on empirical data about the behavior of internal states of a program. The experimental design uses the existing environment of a multi-version experiment previously conducted at the NASA Langley Research Center, in which the launch interceptor problem is used as a model. This allows the controlled experimental investigation of versions with well-known single and multiple faults, and the availability of an oracle permits the determination of the error detection performance of the test. Fault interaction phenomena are observed that have an amplifying effect on the number of error occurrences. Preliminary results indicate that all faults examined so far are detected by the acceptance test. This shows promise for further investigations, and for the employment of this test method on other applications.

Protzel, Peter W.

Low-frequency source parameters of twelve large earthquakes

A global survey of the low-frequency (1-21 mHz) source characteristics of large events are studied. We are particularly interested in events unusually enriched in low-frequency and in events with a short-term precursor. We model the source time function of 12 large earthquakes using teleseismic data at low frequency. For each event we retrieve the source amplitude spectrum in the frequency range between 1 and 21 mHz with the Silver and Jordan method and the phase-shift spectrum in the frequency range between 1 and 11 mHz with the Riedesel and Jordan method. We then model the source time function by fitting the two spectra. Two of these events, the 1980 Irpinia, Italy, and the 1983 Akita-Oki, Japan, are shallow-depth complex events that took place on multiple faults. In both cases the source time function has a length of about 100 seconds. By comparison Westaway and Jackson find 45 seconds for the Irpinia event and Houston and Kanamori about 50 seconds for the Akita-Oki earthquake. The three deep events and four of the seven intermediate-depth events are fast rupturing earthquakes. A single pulse is sufficient to model the source spectra in the frequency range of our interest. Two other intermediate-depth events have slower rupturing processes, characterized by a continuous energy release lasting for about 40 seconds. The last event is the intermediate-depth 1983 Peru-Ecuador earthquake. It was first recognized as a precursive event by Jordan. We model it with a smooth rupturing process starting about 2 minutes before the high frequency origin time superimposed to an impulsive source.

Harabaglia, Paolo

Implementation issues for a compact 6 degree of freedom force reflecting handcontroller with cueing of modes

Teleoperated control requires a master human interface device that can provide haptic input and output which reflect the responses of a slave robotic system. The effort reported in this paper addresses the design and prototyping of a six degree-of-freedom (DOF) Cartesian coordinate hand controller for this purpose. The device design recommended is an XYZ stage attached to a three-roll wrist which positions a flight-type handgrip. Six degrees of freedom are transduced and control brushless DC motor servo electronics similar in design to those used in computer controlled robotic manipulators. This general approach supports scaled force, velocity, and position feedback to aid an operator in achieving telepresence. The generality of the device and control system characteristics allow the use of inverse dynamics robotic control methodology to project slave robot system forces and inertias to the operator (in scaled form) and at the same time to reduce the apparent inertia of the robotic handcontroller itself. The current control design, which is not multiple fault tolerant, can be extended to make flight control or space use possible. The proposed handcontroller will have advantages in space-based applications where an operator must control several robot arms in a simultaneous and coordinated fashion. It will also have applications in intravehicular activities (within the Space Station) such as microgravity experiments in metallurgy and biological experiments that require isolation from the astronauts' environment. For ground applications, the handcontroller will be useful in underwater activities where the generality of the proposed handcontroller becomes an asset for operation of many different manipulator types. Also applications will emerge in the Military, Construction, and Maintenance/Manufacturing areas including ordnance handling, mine removal, NBC (Nuclear, Chemical, Biological) operations, control of vehicles, and operating strength and agility enhanced machines. Future avionics applications including advanced helicopter and aircraft control may also become important.

Jacobus, Heidi

A Unique Power System For The ISS Fluids And Combustion Facility

Unique power control technology has been incorporated into an electrical power control unit (EPCU) for the Fluids and Combustion Facility (FCF). The objective is to maximize science throughput by providing a flexible power system that is easily reconfigured by the science payload. Electrical power is at a premium on the International Space Station (ISS). The EPCU utilizes advanced power management techniques to maximize the power available to the FCF experiments. The EPCU architecture enables dynamic allocation of power from two ISS power channels for experiments. Because of the unique flexible remote power controller (FRPC) design, power channels can be paralleled while maintaining balanced load sharing between the channels. With an integrated and redundant architecture, the EPCU can tolerate multiple faults and still maintain FCF operation. It is important to take full advantage of the EPCU functionality. The EPCU acts as a buffer between the experimenter and the ISS power system with all its complex requirements. However, FCF science payload developers will still need to follow guidelines when designing the FCF payload power system. This is necessary to ensure power system stability, fault coordination, electromagnetic compatibility, and maximum use of available power for gathering scientific data.

Fox, David A.