Search NASASearch

SEARCH · Search NASA

Results for “multiple faults”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Diagnosing faults in autonomous robot plan execution

A major requirement for an autonomous robot is the capability to diagnose faults during plan execution in an uncertain environment. Many diagnostic researches concentrate only on hardware failures within an autonomous robot. Taking a different approach, the implementation of a Telerobot Diagnostic System that addresses, in addition to the hardware failures, failures caused by unexpected event changes in the environment or failures due to plan errors, is described. One feature of the system is the utilization of task-plan knowledge and context information to deduce fault symptoms. This forward deduction provides valuable information on past activities and the current expectations of a robotic event, both of which can guide the plan-execution inference process. The inference process adopts a model-based technique to recreate the plan-execution process and to confirm fault-source hypotheses. This technique allows the system to diagnose multiple faults due to either unexpected plan failures or hardware errors. This research initiates a major effort to investigate relationships between hardware faults and plan errors, relationships which were not addressed in the past. The results of this research will provide a clear understanding of how to generate a better task planner for an autonomous robot and how to recover the robot from faults in a critical environment.

Lam, Raymond K.

Diagnosing faults in autonomous robot plan execution

A major requirement for an autonomous robot is the capability to diagnose faults during plan execution in an uncertain environment. Many diagnostic researches concentrate only on hardware failures within an autonomous robot. Taking a different approach, the implementation of a Telerobot Diagnostic System that addresses, in addition to the hardware failures, failures caused by unexpected event changes in the environment or failures due to plan errors, is described. One feature of the system is the utilization of task-plan knowledge and context information to deduce fault symptoms. This forward deduction provides valuable information on past activities and the current expectations of a robotic event, both of which can guide the plan-execution inference process. The inference process adopts a model-based technique to recreate the plan-execution process and to confirm fault-source hypotheses. This technique allows the system to diagnose multiple faults due to either unexpected plan failures or hardware errors. This research initiates a major effort to investigate relationships between hardware faults and plan errors, relationships which were not addressed in the past. The results of this research will provide a clear understanding of how to generate a better task planner for an autonomous robot and how to recover the robot from faults in a critical environment.

Lam, Raymond K.

Multiple strike slip faults sets: A case study from the Dead Sea transform

In many strike slip tectonic settings, large rotations of crust blocks about vertical axes have been inferred from paleomagnetic data. These blocks are bounded by sets of parallel faults which presumably accommodate the relative motion between the blocks as regional deformation progress. A mechanical model by Nur et al., (1986) suggests that rotations greater than phi sub c equals 25 to 45 degrees must be accommodated by more than one set of faults, with angle phi sub c between their direction; consequently the sum of the angles between sets must be roughly equal to the total tectonic material rotation. To test this model, the authors investigated the fault geometry and field relation of fault sets in the Mt. Hermon area in northern Israel, where paleomagnetic declination implies data 69 degrees plus or minus 13 degrees counter-clockwise block rotation. The statistical and field relation analysis of over 315 faults shows that the faulting is predominantly right lateral strike slip consisting of three distinct sets. The oldest set strikes 253 degrees, the second oldest set strikes 293 degrees and the youngest strikes 339 degrees. This last direction is consistent also with the current north-south direction of the maximum principle stress axis. The angle phi sub c between the first and second sets is 39 degrees and between the second and third sets 46 degrees, in good agreement with the phi sub c angle predicted from mechanical considerations. The sum of the two angles is 85 degrees, in good agreement with the 69 degrees plus or minus 13 degrees CCW paleomagnetically derived rotation. The results suggest specifically that the sequential development of multiple intersecting fault sets is responsible for the faulting in the Mt. Hermon area; and generally that the model of block rotation with multiple faults provides very good simple rules for analyzing very complex fault patterns.

Ron, Hagai

Integrated analysis of error detection and recovery

An integrated modeling and analysis of error detection and recovery is presented. When fault latency and/or error latency exist, the system may suffer from multiple faults or error propagations which seriously deteriorate the fault-tolerant capability. Several detection models that enable analysis of the effect of detection mechanisms on the subsequent error handling operations and the overall system reliability were developed. Following detection of the faulty unit and reconfiguration of the system, the contaminated processes or tasks have to be recovered. The strategies of error recovery employed depend on the detection mechanisms and the available redundancy. Several recovery methods including the rollback recovery are considered. The recovery overhead is evaluated as an index of the capabilities of the detection and reconfiguration mechanisms.

Shin, K. G.

Software for Fault-Tolerant Matrix Multiplication

Formal Linear Algebra Recovery Environment is a computer program for high-performance, fault-tolerant matrix multiplication. The program is based on an extension of the prior theory and practice of fault-tolerant matrix matrix multiplication of the form C = AB. This extension provides low-overhead methods for detecting errors, not only in C, but also in A and/or B. These methods enable the detection of all errors as long as, in a given case, only one entry in A, B, or C is corrupted. The program also provides for following a low-overhead rollback approach to correct errors once detected. Results of computational experiments have demonstrated that the methods implemented in this program work well in practice while imposing an acceptably low level of overhead, relative to high-performance matrix-multiplication methods that do not afford fault tolerance.

Katz, Daniel

The software-implemented fault tolerance /SIFT/ approach to fault tolerant computing

SIFT is an experimental computer designed for highly reliable flight-control service in advanced air transports. Its development was intended to integrate and demonstrate the latest techniques in fault-tolerant computing. During its development, several new problems of some generality were uncovered and solved. The technology developed for the validation of its design is seen as being perhaps as important as the design itself. The SIFT design is described, as is the way in which the design and its validation were shaped by the requirements of its intended application. Attention is also given to reliability and fault tolerance. The most significant feature of the hardware design is the absence of elements that can generate multiple faults, such as shared clocks or data buses. It is noted that the software is realized in only 800 lines of code, of which 80% are in a high-level language.

Goldberg, J.

Stress field rotation or block rotation: An example from the Lake Mead fault system

The Coulomb criterion, as applied by Anderson (1951), has been widely used as the basis for inferring paleostresses from in situ fault slip data, assuming that faults are optimally oriented relative to the tectonic stress direction. Consequently if stress direction is fixed during deformation so must be the faults. Freund (1974) has shown that faults, when arranged in sets, must generally rotate as they slip. Nur et al., (1986) showed how sufficiently large rotations require the development of new sets of faults which are more favorably oriented to the principal direction of stress. This leads to the appearance of multiple fault sets in which older faults are offset by younger ones, both having the same sense of slip. Consequently correct paleostress analysis must include the possible effect of fault and material rotation, in addition to stress field rotation. The combined effects of stress field rotation and material rotation were investigated in the Lake Meade Fault System (LMFS) especially in the Hoover Dam area. Fault inversion results imply an apparent 60 degrees clockwise (CW) rotation of the stress field since mid-Miocene time. In contrast structural data from the rest of the Great Basin suggest only a 30 degrees CW stress field rotation. By incorporating paleomagnetic and seismic evidence, the 30 degrees discrepancy can be neatly resolved. Based on paleomagnetic declination anomalies, it is inferred that slip on NW trending right lateral faults caused a local 30 degrees counter-clockwise (CCW) rotation of blocks and faults in the Lake Mead area. Consequently the inferred 60 degrees CW rotation of the stress field in the LMFS consists of an actual 30 degrees CW rotation of the stress field (as for the entire Great Basin) plus a local 30 degrees CCW material rotation of the LMFS fault blocks.

Ron, Hagai

Fault Slip and Fluid Flow: Seismic Source Analysis to Assess Role of Multiple Slip Patches in Fault Permeability

The relationship between fault reactivation, microearthquakes (MEQs), and permeability evolution during fluid injection plays a critical role in energy harvesting and waste disposal. Recent studies have demonstrated the possibility of predicting fault permeability using cumulative seismic moments of MEQs quantitatively. To understand the underlying physical processes, we conduct fault reactivation experiments using Utah FORGE granitoid and analyze acoustic emission (AE) signals generated during stepwise increases in fluid injection pressure. Frequency analysis of thousands of calibrated AE signals reveals that fault reactivation produces multiple AE source patches with millimeter-scale radii—smaller than the sample fault radius. The cumulative area of the reactivated patches covers the fault multiple times over (∼10x–50x area) for each pressure step. These findings provide mechanistic insight that measured permeability enhancement is not driven by a single large slip event, but by the sequential and interacting activation of multiple slip patches that create a continuous flow pathway.

Nurshal, M. E. M. [Pennsylvania State University,

Tectonic motion in the western United States inferred from very long baseline interferometry measurements, 1980-1986

Over six years of mobile very long baseline interferometry (VLBI) baseline measurements between 12 sites in the western U.S. were used to infer their velocities relative to the North American plate. These velocities were found to be generally consistent with those determined from geologic data and contemporaneous satellite laser ranging measurements in the same region. The discrepancy between the largest velocities determined from the VLBI measurements of 40-48 mm/yr and the relative plate velocity of 50-56 mm/yr predicted from plate motion models is found to be consistent with a broadened distribution of interseismic strain from cyclic activity on the San Andreas and subsidiary faults. The VLBI data are best explained by a cumulative rate of strike-slip motion near the plate boundary of approximately 48 mm/yr, although exclusion of competing values of 56 and 41 mm/yr is based upon very few data. The rates of offshore fault slip inferred from this study range from about 15 mm/yr in central California to negligible amounts in the San Francisco region. Finite element calculations of multiple fault strain distributions show good agreement with systematic variations in the distribution of shear strain along the San Andreas system, as revealed by previous geodetic measurements.

Kroger, Peter M.

Computing Fault Displacements from Surface Deformations

Simplex is a computer program that calculates locations and displacements of subterranean faults from data on Earth-surface deformations. The calculation involves inversion of a forward model (given a point source representing a fault, a forward model calculates the surface deformations) for displacements, and strains caused by a fault located in isotropic, elastic half-space. The inversion involves the use of nonlinear, multiparameter estimation techniques. The input surface-deformation data can be in multiple formats, with absolute or differential positioning. The input data can be derived from multiple sources, including interferometric synthetic-aperture radar, the Global Positioning System, and strain meters. Parameters can be constrained or free. Estimates can be calculated for single or multiple faults. Estimates of parameters are accompanied by reports of their covariances and uncertainties. Simplex has been tested extensively against forward models and against other means of inverting geodetic data and seismic observations. This work

Lyzenga, Gregory

Real-Time Distributed Embedded Oscillator Operating Frequency Monitoring

A document discusses the utilization of embedded clocks inside of operating network data links as an auxiliary clock source to satisfy local oscillator monitoring requirements. Modem network interfaces, typically serial network links, often contain embedded clocking information of very tight precision to recover data from the link. This embedded clocking data can be utilized by the receiving device to monitor the local oscillator for tolerance to required specifications, often important in high-integrity fault-tolerant applications. A device can utilize a received embedded clock to determine if the local or the remote device is out of tolerance by using a single link. The local device can determine if it is failing, assuming a single fault model, with two or more active links. Network fabric components, containing many operational links, can potentially determine faulty remote or local devices in the presence of multiple faults. Two methods of implementation are described. In one method, a recovered clock can be directly used to monitor the local clock as a direct replacement of an external local oscillator. This scheme is consistent with a general clock monitoring function whereby clock sources are clocking two counters and compared over a fixed interval of time. In another method, overflow/underflow conditions can be used to detect clock relationships for monitoring. These network interfaces often provide clock compensation circuitry to allow data to be transferred from the received (network) clock domain to the internal clock domain. This circuit could be modified to detect overflow/underflow conditions of the buffering required and report a fast or slow receive clock, respectively.

Pollock, Julie

Incipient fault detection study for advanced spacecraft systems

A feasibility study to investigate the application of vibration monitoring to the rotating machinery of planned NASA advanced spacecraft components is described. Factors investigated include: (1) special problems associated with small, high RPM machines; (2) application across multiple component types; (3) microgravity; (4) multiple fault types; (5) eight different analysis techniques including signature analysis, high frequency demodulation, cepstrum, clustering, amplitude analysis, and pattern recognition are compared; and (6) small sample statistical analysis is used to compare performance by computation of probability of detection and false alarm for an ensemble of repeated baseline and faulted tests. Both detection and classification performance are quantified. Vibration monitoring is shown to be an effective means of detecting the most important problem types for small, high RPM fans and pumps typical of those planned for the advanced spacecraft. A preliminary monitoring system design and implementation plan is presented.

Milner, G. Martin

The Soil Moisture Acttive Passive Mission: Fault Protection Performance and Lessons Learned

Fault protection as a discipline involves a collection of flight software logic and operational processes for detecting unacceptable anomalous behavior, responding prior to reaching criticality, restricting the propagation of a failure beyond a fault containment region, and recovering the vehicle back to full or degraded functionality if possible. The System Fault Protection (SFP) design for the SMAP Earth orbiter was put to the test during its 90-day vehicle commissioning activities. During this time, the SFP software autonomously protected the vehicle from multiple faults to critical hardware, and the operations team successfully returned the observatory to its science state. The SFP also performed well in the presence of anomalous behavior below true safety limits by not taking unnecessary response actions, instead allowing the operations team time to monitor the behavior. Certain aspects of the SFP design were modified during operations via both parameter updates and a full flight software update in order to better match the vehicle behavior in the flight environment. An evaluation of the SMAP SFP performance during vehicle Commissioning will be provided in this paper, as well as a set of lessons learned largely focused on visibility, SFP mutability in operations, responses to peripheral device faults, and Safe Mode recovery and design. By capturing some of the knowledge gained during SMAP Commissioning, it is intended that this paper provide guidance for making future System Fault Protection designs more robust and supportive of operations.

Clark, Jessica

Model-Based Diagnostics for Propellant Loading Systems

The loading of spacecraft propellants is a complex, risky operation. Therefore, diagnostic solutions are necessary to quickly identify when a fault occurs, so that recovery actions can be taken or an abort procedure can be initiated. Model-based diagnosis solutions, established using an in-depth analysis and understanding of the underlying physical processes, offer the advanced capability to quickly detect and isolate faults, identify their severity, and predict their effects on system performance. We develop a physics-based model of a cryogenic propellant loading system, which describes the complex dynamics of liquid hydrogen filling from a storage tank to an external vehicle tank, as well as the influence of different faults on this process. The model takes into account the main physical processes such as highly nonequilibrium condensation and evaporation of the hydrogen vapor, pressurization, and also the dynamics of liquid hydrogen and vapor flows inside the system in the presence of helium gas. Since the model incorporates multiple faults in the system, it provides a suitable framework for model-based diagnostics and prognostics algorithms. Using this model, we analyze the effects of faults on the system, derive symbolic fault signatures for the purposes of fault isolation, and perform fault identification using a particle filter approach. We demonstrate the detection, isolation, and identification of a number of faults using simulation-based experiments.

Daigle, Matthew John

Airborne Advanced Reconfigurable Computer System (ARCS)

A digital computer subsystem fault-tolerant concept was defined, and the potential benefits and costs of such a subsystem were assessed when used as the central element of a new transport's flight control system. The derived advanced reconfigurable computer system (ARCS) is a triple-redundant computer subsystem that automatically reconfigures, under multiple fault conditions, from triplex to duplex to simplex operation, with redundancy recovery if the fault condition is transient. The study included criteria development covering factors at the aircraft's operation level that would influence the design of a fault-tolerant system for commercial airline use. A new reliability analysis tool was developed for evaluating redundant, fault-tolerant system availability and survivability; and a stringent digital system software design methodology was used to achieve design/implementation visibility.

Bjurman, B. E.

Error detection process - Model, design, and its impact on computer performance

An analytical model is developed for computer error detection processes and applied to estimate their influence on system performance. Faults in the hardware, not in the design, are assumed to be the potential cause of transition to erroneous states during normal operations. The classification properties and associated recovery methods of error detection are discussed. The probability of obtaining an unreliable result is evaluated, along with the resulting computational loss. Error detection during design is considered and a feasible design space is outlined. Extension of the methods to account for the effects of extant multiple faults is indicated.

Shin, K. G.

Simulation Results of a Thermal Power Dispatch System from a Generic Pressurized Water Reactor in Normal and Abnormal Operating Conditions

Amid economic pressures in the U.S. electricity market, nuclear utilities are exploring new revenue streams, including hydrogen production. A generic pressurized water reactor simulator was modified to incorporate a novel design for a TPD system coupled to a hydrogen production plant. Standard malfunctions were included in the simulation design, including steam line breaks at various system locations and flow interruptions in the hydrogen plant due to multiple faults, reflecting anticipated operational challenges. It is imperative that the TPD system operation has a minimal effect on the reactor power, primary coolant system, and turbine system operation and performance. Due to the specific design and application of this TPD system, with the proposed turbine control system changes, the overall impact on the existing plant systems is low. Normal TPD operating scenarios resulted in minor effects on the existing plant systems: reactor power changes by at most 0.2%, and gross generator output changes by 20.5 MWe from 100 MWt of TPD. The most severe malfunction analyzed in this work is a full TPD steam line break downstream of the extraction location, which results in an increase in reactor power of about 0.5%. The gross generator output decreases by 36 MWe, a total decrease of 60 MWe from the full power steady state (FPSS) condition. These results indicate that an industrial hydrogen production plant could be coupled thermally to a nuclear power plant with limited effects on the existing system operation and safety.

08 HYDROGEN

Automatically generated acceptance test: A software reliability experiment

This study presents results of a software reliability experiment investigating the feasibility of a new error detection method. The method can be used as an acceptance test and is solely based on empirical data about the behavior of internal states of a program. The experimental design uses the existing environment of a multi-version experiment previously conducted at the NASA Langley Research Center, in which the launch interceptor problem is used as a model. This allows the controlled experimental investigation of versions with well-known single and multiple faults, and the availability of an oracle permits the determination of the error detection performance of the test. Fault interaction phenomena are observed that have an amplifying effect on the number of error occurrences. Preliminary results indicate that all faults examined so far are detected by the acceptance test. This shows promise for further investigations, and for the employment of this test method on other applications.

Protzel, Peter W.