Search NASA⌕ Search

SEARCH · Search NASA

Results for “System Unreliability”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Neural correspondence to spectrum of environmental uncertainty in multiple-cue probability judgment system with time delay

Despite state-of-the-art technologies like artificial intelligence, human judgment is critically essential in cooperative systems, such as the multi-agent system (MAS), which collect information among agents based on multiple-cue judgment. Human agents can prevent impaired situational awareness of automated agents by confirming situations under environmental uncertainty. System error caused by uncertainty can result in an unreliable system environment, and this environment affects the human agent, resulting in non-optimal decision-making in MAS. Thus, it is necessary to know how human behavior is changed to capture system reliability under uncertainty. Another issue affecting MAS is time delay, which can delay agent information transfer, resulting in low performance and instability. However, it is difficult to find studies on the influence of time delay on human agents. This study is about understanding the human decision-making process under a specific system reliability environment by uncertainty with time delay. We used concepts of expected and unexpected uncertainty to implement reliability of the system usage environment with three types of time delay conditions: no time delay, regular time delay, and irregular time delay conditions. We used electroencephalogram (EEG) for human cognitive neural mechanisms in multiple-cue judgment systems to understand human decision-making. In the reliability of system usage environment, the unreliable system environment significantly creates less memory load by less utilization of system rules for decision-making. In terms of time delay, delayed information delivery does not significantly affect memory load for decision-making.

cognitive process↗

System applications of the fault tolerant memory

Conventional memory technologies currently employed in aerospace applications contribute at least fifty percent to system unreliability (where the system includes CPU, I/O and memory). A fault tolerant memory performs both error correction and memory replacement at the bit plane level. To determine the effects of system design of using a fault tolerant memory in space applications, analysis was performed to determine tradeable hardware configurations that meet the reliability goals of each program. The candidate configurations, which included redundant elements of the computer system with both conventional and fault tolerant memories, were then traded in terms of selection criteria of cost, weight, volume, and power. These trade studies demonstrated that a fault tolerant memory provided significant advantages in terms of cost, weight, and volume. The memory selected for this analysis was a recently developed five fault tolerant memory.

Murphy, L. J.↗

Design Methods and Practices for Fault Prevention and Management in Spacecraft

Integrated Systems Health Management (ISHM) is intended to become a critical capability for all space, lunar and planetary exploration vehicles and systems at NASA. Monitoring and managing the health state of diverse components, subsystems, and systems is a difficult task that will become more challenging when implemented for long-term, evolving deployments. A key technical challenge will be to ensure that the ISHM technologies are reliable, effective, and low cost, resulting in turn in safe, reliable, and affordable missions. To ensure safety and reliability, ISHM functionality, decisions and knowledge have to be incorporated into the product lifecycle as early as possible, and ISHM must be considered as an essential element of models developed and used in various stages during system design. During early stage design, many decisions and tasks are still open, including sensor and measurement point selection, modeling and model-checking, diagnosis, signature and data fusion schemes, presenting the best opportunity to catch and prevent potential failures and anomalies in a cost-effective way. Using appropriate formal methods during early design, the design teams can systematically explore risks without committing to design decisions too early. However, the nature of ISHM knowledge and data is detailed, relying on high-fidelity, detailed models, whereas the earlier stages of the product lifecycle utilize low-fidelity, high-level models of systems and their functionality. We currently lack the tools and processes necessary for integrating ISHM into the vehicle system/subsystem design. As a result, most existing ISHM-like technologies are retrofits that were done after the system design was completed. It is very expensive, and sometimes futile, to retrofit a system health management capability into existing systems. Last-minute retrofits result in unreliable systems, ineffective solutions, and excessive costs (e.g., Space Shuttle TPS monitoring which was considered only after 110 flights and the Columbia disaster). High false alarm or false negative rates due to substandard implementations hurt the credibility of the ISHM discipline. This paper presents an overview of the current state of ISHM design,and a review of formal design methods to make recommendations about possible approaches to enable the ISHM capabilities to be designed-in at the system-level, from the very beginning of the vehicle design process.

Tumer, Irem Y.↗

An Integrated Approach to Life Cycle Analysis

Life Cycle Analysis (LCA) is the evaluation of the impacts that design decisions have on a system and provides a framework for identifying and evaluating design benefits and burdens associated with the life cycles of space transportation systems from a "cradle-to-grave" approach. Sometimes called life cycle assessment, life cycle approach, or "cradle to grave analysis", it represents a rapidly emerging family of tools and techniques designed to be a decision support methodology and aid in the development of sustainable systems. The implementation of a Life Cycle Analysis can vary and may take many forms; from global system-level uncertainty-centered analysis to the assessment of individualized discriminatory metrics. This paper will focus on a proven LCA methodology developed by the Systems Analysis and Concepts Directorate (SACD) at NASA Langley Research Center to quantify and assess key LCA discriminatory metrics, in particular affordability, reliability, maintainability, and operability. This paper will address issues inherent in Life Cycle Analysis including direct impacts, such as system development cost and crew safety, as well as indirect impacts, which often take the form of coupled metrics (i.e., the cost of system unreliability). Since LCA deals with the analysis of space vehicle system conceptual designs, it is imperative to stress that the goal of LCA is not to arrive at the answer but, rather, to provide important inputs to a broader strategic planning process, allowing the managers to make risk-informed decisions, and increase the likelihood of meeting mission success criteria.

Chytka, T. M.↗

Procedure for Failure Mode, Effects, and Criticality Analysis (FMECA)

This document provides guidelines for the accomplishment of Failure Mode, Effects, and Criticality Analysis (FMECA) on the Apollo program. It is a procedure for analysis of hardware items to determine those items contributing most to system unreliability and crew safety problems.

Source record↗

An airline study of advanced technology requirements for advanced high speed commercial transport engines. 2: Engine preliminary design assessment

The advanced technology requirements for an advanced high speed commercial transport engine are presented. The results of the phase 2 study effort cover the following areas: (1) general review of preliminary engine designs suggested for a future aircraft, (2) presentation of a long range view of airline propulsion system objectives and the research programs in noise, pollution, and design which must be undertaken to achieve the goals presented, (3) review of the impact of propulsion system unreliability and unscheduled maintenance on cost of operation, (4) discussion of the reliability and maintainability requirements and guarantees for future engines.

Sallee, G. P.↗

The cost of software fault tolerance

The proposed use of software fault tolerance techniques as a means of reducing software costs in avionics and as a means of addressing the issue of system unreliability due to faults in software is examined. A model is developed to provide a view of the relationships among cost, redundancy, and reliability which suggests strategies for software development and maintenance which are not conventional.

Migneault, G. E.↗

Markov reliability models for digital flight control systems

The reliability of digital flight control systems can often be accurately predicted using Markov chain models. The cost of numerical solution depends on a model's size and stiffness. Acyclic Markov models, a useful special case, are particularly amenable to efficient numerical solution. Even in the general case, instantaneous coverage approximation allows the reduction of some cyclic models to more readily solvable acyclic models. After considering the solution of single-phase models, the discussion is extended to phased-mission models. Phased-mission reliability models are classified based on the state restoration behavior that occurs between mission phases. As an economical approach for the solution of such models, the mean failure rate solution method is introduced. A numerical example is used to show the influence of fault-model parameters and interphase behavior on system unreliability.

Mcgough, John↗

System Study: Emergency Power System 1998-2022

This report presents an unreliability evaluation of the emergency power system (EPS) at 93 U.S. commercial operating nuclear reactors. New Standardized Plant Analysis Risk (SPAR) models with the most recent SPAR parameter update results were used in this report. Demand, run hour, and failure data from 1998–2022 for selected components were obtained from the Institute of Nuclear Power Operations Industry Reporting and Information System. The unreliability results are trended for the most recent 10 year period while yearly estimates for system unreliability are provided for the entire active period. No statistically significant increasing or decreasing trends were identified in the industry-wide estimates of EPS system start-only unreliability, but a highly statistically significant decreasing trend was identified in the industry-wide estimates of EPS system 8-hour mission unreliability.

99 GENERAL AND MISCELLANEOUS↗

System Study: Emergency Power System 1998-2024

This report presents an unreliability evaluation of the emergency power system (EPS) at 93 U.S. commercial operating nuclear reactors. New Standardized Plant Analysis Risk (SPAR) models with the most recent SPAR parameter update results were used in this report. Demand, run hour, and failure data from 1998 to 2024 for selected components were obtained from the Institute of Nuclear Power Operations Industry Reporting and Information System. The unreliability results are trended for the most recent 10 year period while yearly estimates for system unreliability are provided for the entire active period. No statistically significant increasing or decreasing trends were identified in the industry-wide estimates of EPS system start-only unreliability, but a statistically significant decreasing trend was identified in the industry-wide estimates of EPS system 24-hour mission unreliability.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

System Study: Auxiliary Feedwater 1998-2022

This report presents an unreliability evaluation of the auxiliary feedwater (AFW) system at 62 U.S. commercial operating nuclear reactors. New Standardized Plant Analysis Risk (SPAR) models with the most recent SPAR parameter update results were used in this report. Demand, run hour, and failure data from calendar year 1998–2022 for selected components were obtained from the Institute of Nuclear Power Operations Industry Reporting and Information System. The unreliability results are trended for the most recent 10 year period while yearly estimates for system unreliability are provided for the entire active period. No statistically significant increasing or decreasing trends were identified in the industry-wide estimates of AFW system start-only unreliability, but a highly statistically significant decreasing trend was identified in the industry-wide estimates of AFW system 8-hour mission unreliability.

99 GENERAL AND MISCELLANEOUS↗

Generating synthetic signaling networks for in silico modeling studies

Predictive models of signaling pathways have proven to be difficult to develop. Reasons include the uncertainty in the number of species, the complexity in species’ interactions, and the sparseness and uncertainty in experimental data. Traditional approaches to developing mechanistic models rely on collecting experimental data and fitting a single model to that data. This approach works for simple systems but has proven unreliable for complex systems such as biological signaling networks. For example, uncertainty and sparseness of the data often result in overfitted models that have little predictive value beyond recapitulating the experimental data itself. Thus, there is a need to develop new approaches to create predictive mechanistic models of complex systems. However, to determine the effectiveness of any new algorithm, a baseline model is needed to test its performance. To meet this need, we developed a method for generating artificial synthetic networks that are reasonably realistic and thus can be treated as ground truth models. These synthetic models can then be used to generate synthetic data for developing and testing algorithms designed to recover the underlying network topology and associated parameters. Here, we describe a simple approach for generating synthetic signaling networks that can be used for this purpose.

42 ENGINEERING↗

An Efficient Approach for the Reliability Analysis of Phased-Mission Systems with Dependent Failures

We consider the reliability analysis of phased-mission systems with common-cause failures in this paper. Phased-mission systems (PMS) are systems supporting missions characterized by multiple, consecutive, and nonoverlapping phases of operation. System components may be subject to different stresses as well as different reliability requirements throughout the course of the mission. As a result, component behavior and relationships may need to be modeled differently from phase to phase when performing a system-level reliability analysis. This consideration poses unique challenges to existing analysis methods. The challenges increase when common-cause failures (CCF) are incorporated in the model. CCF are multiple dependent component failures within a system that are a direct result of a shared root cause, such as sabotage, flood, earthquake, power outage, or human errors. It has been shown by many reliability studies that CCF tend to increase a system's joint failure probabilities and thus contribute significantly to the overall unreliability of systems subject to CCF.We propose a separable phase-modular approach to the reliability analysis of phased-mission systems with dependent common-cause failures as one way to meet the above challenges in an efficient and elegant manner. Our methodology is twofold: first, we separate the effects of CCF from the PMS analysis using the total probability theorem and the common-cause event space developed based on the elementary common-causes; next, we apply an efficient phase-modular approach to analyze the reliability of the PMS. The phase-modular approach employs both combinatorial binary decision diagram and Markov-chain solution methods as appropriate. We provide an example of a reliability analysis of a PMS with both static and dynamic phases as well as CCF as an illustration of our proposed approach. The example is based on information extracted from a Mars orbiter project. The reliability model for this orbiter considers the various phases of Launch, Cruise, Mars Orbit Insertion, and Orbit. Some of the CCF for the orbiter in this mission include environmental effects, such as micrometeoroids, human operator errors, and software errors.

reliability analysis↗

System Study: Auxiliary Feedwater 1998-2024

This report presents an unreliability evaluation of the auxiliary feedwater (AFW) system at 62 U.S. commercial operating nuclear reactors. New Standardized Plant Analysis Risk (SPAR) models with the most recent SPAR parameter update results were used in this report. Demand, run hour, and failure data from calendar years 1998 to 2024 for selected components were obtained from the Institute of Nuclear Power Operations Industry Reporting and Information System. The unreliability results are trended for the most recent 10 year period and yearly estimates for system unreliability are provided for the entire active period. No statistically significant increasing or decreasing trends were identified in the industry-wide estimates of AFW system start-only unreliability, but a statistically significant decreasing trend was identified in the industry-wide estimates of AFW system 24-hour mission unreliability.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Algorithms for Multiple Fault Diagnosis With Unreliable Tests

In this paper, we consider the problem of constructing optimal and near-optimal multiple fault diagnosis (MFD) in bipartite systems with unreliable (imperfect) tests. It is known that exact computation of conditional probabilities for multiple fault diagnosis is NP-hard. The novel feature of our diagnostic algorithms is the use of Lagrangian relaxation and subgradient optimization methods to provide: (1) near optimal solutions for the MFD problem, and (2) upper bounds for an optimal branch-and-bound algorithm. The proposed method is illustrated using several examples. Computational results indicate that: (1) our algorithm has superior computational performance to the existing algorithms (approximately three orders of magnitude improvement), (2) the near optimal algorithm generates the most likely candidates with a very high accuracy, and (3) our algorithm can find the most likely candidates in systems with as many as 1000 faults.

Shakeri, Mojdeh↗

System Study: Isolation Condenser 1998-2022

This report presents an unreliability evaluation of the isolation condenser (ISO) system at three U.S. commercial operating boiling water reactors. New Standardized Plant Analysis Risk (SPAR) models with the most recent SPAR parameter update results were used in this report. Demand, run hour, and failure data from calendar year 1998–2022 for selected components were obtained from the Institute of Nuclear Power Operations (INPO) Industry Reporting and Information System (IRIS). The unreliability results are trended for the most recent 10 year period while yearly estimates for system unreliability are provided for the entire active period. No statistically significant increasing or decreasing trends were identified in the ISO results.

99 GENERAL AND MISCELLANEOUS↗

System Study: Reactor Core Isolation Cooling 1998-2022

This report presents an unreliability evaluation of the reactor core isolation cooling (RCIC) system at 28 U.S. commercial operating boiling water reactors. New Standardized Plant Analysis Risk (SPAR) models with the most recent SPAR parameter update results were used in this report. Demand, run hour, and failure data from calendar years 1998–2022 for selected components were obtained from the Institute of Nuclear Power Operations Industry Reporting and Information System. The unreliability results are trended for the most recent 10 year period while yearly estimates for system unreliability are provided for the entire active period. Statistically significant decreasing trends were identified in both the RCIC system start-only unreliability and 8-hour mission unreliability.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗