Search NASA⌕ Search

SEARCH · Search NASA

Results for “Fault Injection”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Portable Health Algorithms Test System

A document discusses the Portable Health Algorithms Test (PHALT) System, which has been designed as a means for evolving the maturity and credibility of algorithms developed to assess the health of aerospace systems. Comprising an integrated hardware-software environment, the PHALT system allows systems health management algorithms to be developed in a graphical programming environment, to be tested and refined using system simulation or test data playback, and to be evaluated in a real-time hardware-in-the-loop mode with a live test article. The integrated hardware and software development environment provides a seamless transition from algorithm development to real-time implementation. The portability of the hardware makes it quick and easy to transport between test facilities. This hard ware/software architecture is flexible enough to support a variety of diagnostic applications and test hardware, and the GUI-based rapid prototyping capability is sufficient to support development execution, and testing of custom diagnostic algorithms. The PHALT operating system supports execution of diagnostic algorithms under real-time constraints. PHALT can perform real-time capture and playback of test rig data with the ability to augment/ modify the data stream (e.g. inject simulated faults). It performs algorithm testing using a variety of data input sources, including real-time data acquisition, test data playback, and system simulations, and also provides system feedback to evaluate closed-loop diagnostic response and mitigation control.

Melcher, Kevin J.↗

A Cryogenic Fluid System Simulation in Support of Integrated Systems Health Management

Simulations serve as important tools throughout the design and operation of engineering systems. In the context of sys-tems health management, simulations serve many uses. For one, the underlying physical models can be used by model-based health management tools to develop diagnostic and prognostic models. These simulations should incorporate both nominal and faulty behavior with the ability to inject various faults into the system. Such simulations can there-fore be used for operator training, for both nominal and faulty situations, as well as for developing and prototyping health management algorithms. In this paper, we describe a methodology for building such simulations. We discuss the design decisions and tools used to build a simulation of a cryogenic fluid test bed, and how it serves as a core technology for systems health management development and maturation.

cryogenics↗

Cryogenic Fuel Valve Testbed Development

The goal for this project is to update the cryogenic valve testbed program in LabVIEW to schedule and automate tests and experiments. By using an automated system, tens or hundreds of tests may be performed. This will ensure that accurate data is being collected for testing of the remaining useful life and end of life predictions. From the data obtained, new diagnostic and prognostic methods will be developed to manage or predict potential leaks which may occur in the future. The Cryogenic valve testbed injects controlled faults into the cryogenic fuel valve system in order to accurately determine failure behavior.

Prognostics↗

Validation of an SEU simulation technique for a complex processor: PowerPC7400

Published data on the processors sensitivites with respect to SEU is generally obtained from radiation ground testing during which the program is executed by the DUT consists in the sequential inspection of each of the processor memory cells accessible to the user, through the execution of a suitable instruction sequence. In such programs, so-called static tests, typically considered memory cells are general-purpose registers, special registers (program counter, stack pointer...) and internal memory. Nevertheless, the register use and duty cycle of the final application will be very different, including using instructions no in the static tests and disturbing other potential SEU targets. The ideal would be the use of the final application program for the radiation ground testing, but generally this program is either unknown or unavailable when the qualification testing is performed on candidate circuits to space projects.

Radiation↗

Uncovering Hazards Using a Multi-Objective Optimization to Explore the Faulty State-Space

Considering resilience when designing complex engineered systems is crucial to ensure the system is safe under unexpected hazardous scenarios. Traditional risk-based approaches, such as Failure Modes and Effects Analysis (FMEA) are useful for designing the system to mitigate hazardous scenarios that can be identified by the designer, but often require experience or prior knowledge of system failures to generate. More recently, researchers have developed simulation tools that enable the designer to model large sets of hazardous scenarios (driven by both internal faults and external factors) through simulation. While these tools enable a wider scope of fault modes to be evaluated (e.g., by injecting combined set of fault modes or injecting modes at different times), the resulting assessments (like FMEA) still require knowledge of the specific modes to be evaluated. However, failure to analyze a wide variety of fault scenarios can lead to an incomplete picture of the system resilience, especially to "surprise events'' which may be difficult for the designer to identify and predict beforehand. To overcome this challenge, previous work developed a fault sampling approach for resilience simulations which would procedurally-generate a wide variety of potential faults by systematically perturbing the health states of the system. While the resulting fault modes generated covered a much larger space hazards than would be otherwise considered (and identified many unique failure trajectories which would not have otherwise been identified), it also significantly increased the computational cost of the analysis and resulted in the simulation and analysis of a large set of essentially duplicate scenarios. Additionally, as the number of dimensions in the faulty state-space increases, the full elaboration of possible modes becomes computationally infeasible, justifying the use of a more targeted search. To resolve this limitation, this work proposes the use of a multiobjective optimization algorithm to search the health state space for potential fault modes that are both (1) hazardous and (2) unique. To solve this type of problem, this work proposes the use of a cooperative co-evolutionary algorithm. To demonstrate this approach, it will be applied to a model of an autonomous rover which uses line markings to navigate, focusing on potential hazards in the drive system which could cause the rover to crash. To determine the merit of the approach, it will further be compared with the previously-presented range elaboration approach and a random mode generation approach on the basis of computational efficiency and found modes.

Resilience↗

Fault-tolerance experiments with the JPL STAR computer.

Results of fault-tolerance experiments performed using an experimental computer with dynamic (standby) redundancy, including replaceable subsystems and a 'program rollback' provision to eliminate transient-caused errors. After a brief review of the specification of fault-tolerance with respect to transient faults, including a description of the method of injection of transient faults in software and system tests, fault-tolerance experiments carried out with this computer with regard to the determination of fault classes, software verification, system verification, and recovery stability are summarized. A test and repair processor is described which constitutes a special monitor unit of the computer and is used to obtain information for fault detection in the other subsystems of the computer and to ensure that proper recovery occurs when a fault is detected.

Avizienis, A.↗

Experimental Validation of Model-Based Prognostics for Pneumatic Valves

Because valves control many critical operations, they are prime candidates for deployment of prognostic algorithms. But, similar to the situation with most other components, examples of failures experienced in the field are hard to come by. This lack of data impacts the ability to test and validate prognostic algorithms. A solution sometimes employed to overcome this shortcoming is to perform run-to-failure experiments in a lab. However, the mean time to failure of valves is typically very high (possibly lasting decades), preventing evaluation within a reasonable time frame. Therefore, a mechanism to observe development of fault signatures considerably faster is sought. Described here is a testbed that addresses these issues by allowing the physical injection of leakage faults (which are the most common fault mode) into pneumatic valves. What makes this testbed stand out is the ability to modulate the magnitude of the fault almost arbitrarily fast. With that, the performance of end-of-life estimation algorithms can be tested. Further, the testbed is mobile and can be connected to valves in the field. This mobility helps to bring the overall process of prognostic algorithm development for this valve a step closer to validation. The paper illustrates the development of a model-based prognostic approach that uses data from the testbed for partial validation.

Chetan S Kulkarni↗

Validation of Model-Based Prognostics for Pneumatic Valves in a Cryogenic Fueling Demonstration Testbed

Because valves control many critical operations, they are prime candidates for deployment of prognostic algorithms. But, similar to the situation with most other components, examples of failures experienced in the field are hard to come by. This lack of data impacts the ability to test and validate prognostic algorithms. A solution sometimes employed to overcome this shortcoming is to perform run to failure experiments in a lab. However, the mean time to failure of valves is typically very high (possibly lasting decades), preventing evaluation within a reasonable time frame. Therefore, a mechanism to observe development of fault signatures considerably faster is sought. Described here is a testbed that addresses these issues by allowing the physical injection of leakage faults (which are the most common fault mode) into pneumatic valves. What makes this testbed stand out is the ability to modulate the magnitude of the fault almost arbitrarily fast. With that, the performance of end-of-life estimation algorithms can be tested. Further, the testbed is mobile and can be connected to valves in the field. This mobility helps to bring the overall process of prognostic algorithm development for this valve a step closer to validation. The paper illustrates the development of a model-based prognostic approach that uses data from the testbed for partial validation.

Kulkarni, Chetan S.↗

FOCUS - An experimental environment for fault sensitivity analysis

FOCUS, a simulation environment for conducting fault-sensitivity analysis of chip-level designs, is described. The environment can be used to evaluate alternative design tactics at an early design stage. A range of user specified faults is automatically injected at runtime, and their propagation to the chip I/O pins is measured through the gate and higher levels. A number of techniques for fault-sensitivity analysis are proposed and implemented in the FOCUS environment. These include transient impact assessment on latch, pin and functional errors, external pin error distribution due to in-chip transients, charge-level sensitivity analysis, and error propagation models to depict the dynamic behavior of latch errors. A case study of the impact of transient faults on a microprocessor-based jet-engine controller is used to identify the critical fault propagation paths, the module most sensitive to fault propagation, and the module with the highest potential for causing external errors.

Choi, Gwan S.↗

Design analysis and computer-aided performance evaluation of shuttle orbiter electrical power system. Volume 1: Summary

Studies were conducted to develop appropriate space shuttle electrical power distribution and control (EPDC) subsystem simulation models and to apply the computer simulations to systems analysis of the EPDC. A previously developed software program (SYSTID) was adapted for this purpose. The following objectives were attained: (1) significant enhancement of the SYSTID time domain simulation software, (2) generation of functionally useful shuttle EPDC element models, and (3) illustrative simulation results in the analysis of EPDC performance, under the conditions of fault, current pulse injection due to lightning, and circuit protection sizing and reaction times.

Source record↗

Error latency measurements in symbolic architectures

Error latency, the time that elapses between the occurrence of an error and its detection, has a significant effect on reliability. In computer systems, failure rates can be elevated during a burst of system activity due to increased detection of latent errors. A hybrid monitoring environment is developed to measure the error latency distribution of errors occurring in main memory. The objective of this study is to develop a methodology for gauging the dependability of individual data categories within a real-time application. The hybrid monitoring technique is novel in that it selects and categorizes a specific subset of the available blocks of memory to monitor. The precise times of reads and writes are collected, so no actual faults need be injected. Unlike previous monitoring studies that rely on a periodic sampling approach or on statistical approximation, this new approach permits continuous monitoring of referencing activity and precise measurement of error latency.

Young, L. T.↗

A Testbed for Evaluating Lunar Habitat Autonomy Architectures

A lunar outpost will involve a habitat with an integrated set of hardware and software that will maintain a safe environment for human activities. There is a desire for a paradigm shift whereby crew will be the primary mission operators, not ground controllers. There will also be significant periods when the outpost is uncrewed. This will require that significant automation software be resident in the habitat to maintain all system functions and respond to faults. JSC is developing a testbed to allow for early testing and evaluation of different autonomy architectures. This will allow evaluation of different software configurations in order to: 1) understand different operational concepts; 2) assess the impact of failures and perturbations on the system; and 3) mitigate software and hardware integration risks. The testbed will provide an environment in which habitat hardware simulations can interact with autonomous control software. Faults can be injected into the simulations and different mission scenarios can be scripted. The testbed allows for logging, replaying and re-initializing mission scenarios. An initial testbed configuration has been developed by combining an existing life support simulation and an existing simulation of the space station power distribution system. Results from this initial configuration will be presented along with suggested requirements and designs for the incremental development of a more sophisticated lunar habitat testbed.

Lawler, Dennis G.↗

Integrated Software Health Management for Aircraft GN and C

Modern aircraft rely heavily on dependable operation of many safety-critical software components. Despite careful design, verification and validation (V&V), on-board software can fail with disastrous consequences if it encounters problematic software/hardware interaction or must operate in an unexpected environment. We are using a Bayesian approach to monitor the software and its behavior during operation and provide up-to-date information about the health of the software and its components. The powerful reasoning mechanism provided by our model-based Bayesian approach makes reliable diagnosis of the root causes possible and minimizes the number of false alarms. Compilation of the Bayesian model into compact arithmetic circuits makes SWHM feasible even on platforms with limited CPU power. We show initial results of SWHM on a small simulator of an embedded aircraft software system, where software and sensor faults can be injected.

Schumann, Johann↗

Demonstration of Prognostics-Enabled Decision Making Algorithms on a Hardware Mobile Robot Test Platform

Prognostics-enabled Decision Making (PDM) is an emerging research area that aims to integrate prognostic health information and knowledge about the future operating conditions into the process of selecting subsequent actions for the system. Previous work developing and testing PDM algorithms has been done in simulation; this paper describes the effort leading to a successful demonstration of PDM algorithms on a hardware mobile robot platform. The hardware platform, based on the K11 planetary rover prototype, was modified to allow injection of selected fault modes related to the rover’s electrical power subsystem. The PDM algorithms were adapted to the hardware platform, including development of a software module framework, a new route planner, and modifications to increase the algorithms’ robustness to sensor noise and system timing issues. A set of test scenarios was chosen to demonstrate the algorithms’ capabilities. The modifications to run with a hardware platform, the test scenarios, and the test results are described in detail. The results show a successful use of PDM algorithms on a hardware test platform to optimize mission planning in the presence of electrical system faults.

Prognosis↗

Detecting Latent Faults In Digital Flight Controls

Report discusses theory, conduct, and results of tests involving deliberate injection of low-level faults into digital flight-control system. Part of study of effectiveness of techniques for detection of and recovery from faults, based on statistical assessment of inputs and outputs of parts of control systems. Offers exceptional new capability to establish reliabilities of critical digital electronic systems in aircraft.

Mcgough, John↗

Test vectors development and optimization for a microprocessor

This paper describes a method for generating and optimizing test vectors for a microprocessor, with the aid of a fault simulator implemented entirely by hardware. The development and optimization of test vectors has been done on a tester, with the fault simulator plugged directly into the test head. The fault simulator is capable of automatically injecting over a thousand single or multiple stuck faults in the sequential and combinatorial parts of the microprocessor. The test vectors developed by a programmer working interactively with the tester were applied through the tester to the fault simulator, and the percent of faults detected was measured. The vectors were developed and optimized for the 1802 microprocessor, with the objective of detecting 100% of the single stuck faults with a minimum set of vectors. Experimental results show that 99.7% of the single stuck faults are being detected with approximately 14,000 vectors.

Timoc, C. C.↗

Lightning Pin Injection Test: MOSFETS in "ON" State

The test objective was to evaluate MOSFETs for induced fault modes caused by pin-injecting a standard lightning waveform into them while operating. Lightning Pin-Injection testing was performed at NASA LaRC. Subsequent fault-mode and aging studies were performed by NASA ARC researchers using the Aging and Characterization Platform for semiconductor components. This report documents the test process and results, to provide a basis for subsequent lightning tests. The ultimate IVHM goal is to apply prognostic and health management algorithms using the features extracted during aging to allow calculation of expected remaining useful life. A survey of damage assessment techniques based upon inspection is provided, and includes data for optical microscope and X-ray inspection. Preliminary damage assessments based upon electrical parameters are also provided.

Ely, Jay J.↗

Analytical Redundancy Using Kalman Filters for Rocket Engine Sensor Validation

The use of sensor redundancy is crucial in aerospace systems to maintain safe, reliable operation. While hardware redundancy is more common in application, analytical redundancy can provide a viable alternative in systems where the installation of multiple redundant sensors is not viable. To this end, the use of Kalman filters to analytically validate sensor measurements within rocket engines was explored. First, a dynamic model of the RS 25 engine, a derivative of the Space Shuttle Main Engine (SSME), was reduced to a subset of relations, focused around the main combustion chamber pressure. These relations were used within the Kalman filter algorithm to generate an estimate of sensor measurements to be compared with true measurements for data validation purposes. By using a bank of Kalman filters, the residuals between the estimated and true measurements were used to detect and isolate sensor faults. Through fault simulations, the sensor validation performance of this Kalman filter bank design was compared to a hardware redundancy check. Sensor bias and drift faults of various magnitudes were injected into nominal RS 25 engine test data. Results for both approaches show comparable fault detection with most bias faults found nearly instantaneously by both algorithms. Drift fault detection results show certain cases where one algorithm is faster than the other. The key advantage of the Kalman filter algorithm is shown in fault isolation performance where it can isolate faults between two redundant sensors while the hardware redundancy comparisons cannot.

sensors↗