Search NASA⌕ Search

SEARCH · Search NASA

Results for “Reliability analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 271 records · Page 15

Design for Reliability (DfR) in Space Life Support

The engineering process of Design for Reliability (DfR) is well established in the automotive and aerospace industries. DfR should be useful in the future development of space life support systems. DfR is a sequence of tasks that develop system requirements and plan reliability analysis and testing. First and fundamentally, the reliability requirement is defined. Next the system reliability model is developed, often using a reliability block diagram. The overall system reliability requirement is allocated to the subsystems and an estimate of the attainable reliability is made. This expected reliability can be improved by simplifying the design by removing components or by replacing less reliable components. Improving reliability can require difficult compromises, such as reducing performance requirements, increasing budget, or extending testing. The actual system reliability can be determined only by testing, which should continue long enough to provide the required confidence in the measured value. New systems often have unexpected design errors that cause failures in early testing. The usual reliability improvement process of testing, finding the failure modes, and redesigning to remove them reduces the failure rate and is referred to as “reliability growth.” After redesign has been completed, the system should be further tested to determine the actual achieved reliability more accurately. If the final system failure rate is too high, redundant systems can be used to improve overall operational reliability. Adding redundancy simply to increase the one- or two-fault tolerance metric may sometimes reduce reliability. Reliability can be improved in three ways: redesigning the system to include more reliable subsystems and components, reliability growth testing and failure mode removal, and by using parallel redundant systems. DfR should combine these approaches to achieve the required reliability while managing performance, cost, and schedule.

Reliability↗

Determining Component Probability using Problem Report Data for Ground Systems used in Manned Space Flight

During the shuttle era NASA utilized a failure reporting system called the Problem Reporting and Corrective Action (PRACA) it purpose was to identify and track system non-conformance. The PRACA system over the years evolved from a relatively nominal way to identify system problems to a very complex tracking and report generating data base. The PRACA system became the primary method to categorize any and all anomalies from corrosion to catastrophic failure. The systems documented in the PRACA system range from flight hardware to ground or facility support equipment. While the PRACA system is complex, it does possess all the failure modes, times of occurrence, length of system delay, parts repaired or replaced, and corrective action performed. The difficulty is mining the data then to utilize that data in order to estimate component, Line Replaceable Unit (LRU), and system reliability analysis metrics. In this paper, we identify a methodology to categorize qualitative data from the ground system PRACA data base for common ground or facility support equipment. Then utilizing a heuristic developed for review of the PRACA data determine what reports identify a credible failure. These data are the used to determine inter-arrival times to perform an estimation of a metric for repairable component-or LRU reliability. This analysis is used to determine failure modes of the equipment, determine the probability of the component failure mode, and support various quantitative differing techniques for performing repairable system analysis. The result is that an effective and concise estimate of components used in manned space flight operations. The advantage is the components or LRU's are evaluated in the same environment and condition that occurs during the launch process.

Monaghan, Mark W.↗

Probabilistic finite elements for fracture and fatigue analysis

The fusion of the probabilistic finite element method (PFEM) and reliability analysis for probabilistic fracture mechanics (PFM) is presented. A comprehensive method for determining the probability of fatigue failure for curved crack growth was developed. The criterion for failure or performance function is stated as: the fatigue life of a component must exceed the service life of the component; otherwise failure will occur. An enriched element that has the near-crack-tip singular strain field embedded in the element is used to formulate the equilibrium equation and solve for the stress intensity factors at the crack-tip. Performance and accuracy of the method is demonstrated on a classical mode 1 fatigue problem.

Liu, W. K.↗

Rancor-HUNTER: Using a Simulator Engine for Realistic Human Performance Modeling of Nuclear Power Operations

The Human Unimodel for Nuclear Technology to Enhance Reliability (HUNTER) is a software system to simulate human performance in support of human reliability analysis (HRA) in nuclear power plants. This paper summarizes recent work to integrate HUNTER with a plant simulator, namely the Rancor Microworld Simulator. Rancor is an offshoot of earlier work at Idaho National Laboratory (INL) to support plant modernization. The graphical software tools used to mimic digital human-system interface upgrades at INL’s Human Systems Simulation Laboratory were linked to the Rancor Microworld Simulator, an INL-developed simplified plant model. HUNTER becomes a “virtual operator” coupled to the Rancor simulator, thereby allowing a tight coupling between a digital human twin and a digital twin of the plant. Rancor-HUNTER may be run through Monte Carlo iterations across a dynamic range of performance shaping factors, thereby producing distributions of human performance in terms of procedure paths, errors instantiations, and task durations. This paper overviews the various unique features of Rancor-HUNTER and presents an example run of Rancor-HUNTER for a startup scenario.

99 - GENERAL AND MISCELLANEOUS↗

Estimating the Contributions to Human Error Probability from the Convolution of the Distribution of Time Available and Time Required

As part of their duties, Human Reliability Analysis must often evaluate if crews in nuclear power plants (NPPs) can complete tasks associated with a human-failure event within time limits. For example, the time required in NPP scenarios is determined by systematic and structured walkthroughs, feasibility studies, recorded times from training exercises, and interviews with experienced operators and experts. Typically, a point estimate is derived for the estimate (mean, maximum, or 95th percentile of time required). Using point-estimate values can mask the risk associated with variability among crews, plant conditions and set-up, environmental conditions, and other impact factors under which these actions are executed. While point estimates for time required and time available have served the industry well, without considering the uncertainty they could lead to biased understanding about the risk. The Integrated Human Event Analysis System - General Methodology (IDHEAS-G) model (developed by the US Nuclear Regulatory Commission, NRC) for human error probability calculates human error probability by summing two probabilities: insufficient time and cognitive error. As such, the model takes a more holistic approach by considering the full distributions for time required and time available to calculate the human error probability because the time available to complete the task is insufficient. In this study, we expand on the work of the NRC and discuss methods for estimating these time considerations. For example, for the time required, the impact of Performance Influencing Factors (PIFs) on the distribution was divided into impacts that are aleatory in nature, such as crew-to-crew variability, and those that are epistemic (i.e., the PIFs). Starting with the factors that introduce aleatory uncertainty, a first-order distribution was developed from a large set of time required (i.e., NPP task completion times) data for the range of operator actions that occur in the NPP control room under simulated accident conditions. The first-order distribution can then be adjusted to account for epistemic uncertainty using research associated with the impact of applicable PIFs on the time required. We also develop guidance for analysts to address the probability distributions for the time available. The guidance we developed on how to estimate time required and time available distributions is based on the identification of pertinent research and data, data analyses, and expert knowledge elicitation.

human error probability, human performance, time e↗

Reliability/safety analysis of a fly-by-wire system

An analysis technique has been developed to estimate the reliability of a very complex, safety-critical system by constructing a diagram of the reliability equations for the total system. This diagram has many of the characteristics of a fault-tree or success-path diagram, but is much easier to construct for complex redundant systems. The diagram provides insight into system failure characteristics and identifies the most likely failure modes. A computer program aids in the construction of the diagram and the computation of reliability. Analysis of the NASA F-8 Digital Fly-by-Wire Flight Control System is used to illustrate the technique.

Brock, L. D.↗

Reliability-based failure analysis of brittle materials

The reliability of brittle materials under a generalized state of stress is analyzed using the Batdorf model. The model is modified to include the reduction in shear due to the effect of the compressive stress on the microscopic crack faces. The combined effect of both surface and volume flaws is included. Due to the nature of fracture of brittle materials under compressive loading, the component is modeled as a series system in order to establish bounds on the probability of failure. A computer program was written to determine the probability of failure employing data from a finite element analysis. The analysis showed that for tensile loading a single crack will be the cause of total failure but under compressive loading a series of microscopic cracks must join together to form a dominant crack.

Powers, Lynn M.↗

Human Reliability and the Cost of Doing Business

Most businesses recognize that people will make mistakes and assume errors are just part of the cost of doing business, but does it need to be? Companies with high risk, or major consequences, should consider the effect of human error. In a variety of industries, Human Errors have caused costly failures and workplace injuries. These have included: airline mishaps, medical malpractice, administration of medication and major oil spills have all been blamed on human error. A technique to mitigate or even eliminate some of these costly human errors is the use of Human Reliability Analysis (HRA). Various methodologies are available to perform Human Reliability Assessments that range from identifying the most likely areas for concern to detailed assessments with human error failure probabilities calculated. Which methodology to use would be based on a variety of factors that would include: 1) how people react and act in different industries, and differing expectations based on industries standards, 2) factors that influence how the human errors could occur such as tasks, tools, environment, workplace, support, training and procedure, 3) type and availability of data and 4) how the industry views risk & reliability influences ( types of emergencies, contingencies and routine tasks versus cost based concerns). The Human Reliability Assessments should be the first step to reduce, mitigate or eliminate the costly mistakes or catastrophic failures. Using Human Reliability techniques to identify and classify human error risks allows a company more opportunities to mitigate or eliminate these risks and prevent costly failures.

DeMott, Diana↗

Failure-Tolerant Avionics for Crewed Space Systems Recommended Best Practices

This paper provides an overview of some of the major steps needed to mature and justify the design of an avionics system for crewed spacecraft. It is organized as a collection of artifacts or pieces of evidence that NASA needs to assess the system at design reviews, including a functional failure modes and effects analysis (FFMEA), fault containment region (FCR) definitions, the failure hypothesis, and reliability analysis. This paper is intended as a reference for designers working on NASA crewed spaceflight projects, reliability engineers responsible for avionics system assessments, and program managers wanting to understand what evidence is required at design reviews to ensure crew safety and mission success.

Avionics↗

Selected reliability studies for the NERVA program

An investigation was made into certain methods of reliability analysis that are particularly suitable for complex mechanisms or systems in which there are many interactions. The methods developed were intended to assist in the design of such mechanisms, especially for analysis of failure sensitivity to parameter variations and for estimating reliability where extensive and meaningful life testing is not feasible. The system is modeled by a network of interconnected nodes. Each node is a state or mode of operation, or is an input or output node, and the branches are interactions. The network, with its probabilistic and time-dependent paths is also analyzed for reliability and failure modes by a Monte Carlo, computerized simulation of system performance.

Hoover, S. V.↗

The 747 primary flight control systems reliability and maintenance study

The major operational characteristics of the 747 Primary Flight Control Systems (PFCS) are described. Results of reliability analysis for separate control functions are presented. The analysis makes use of a NASA computer program which calculates reliability of redundant systems. Costs for maintaining the 747 PFCS in airline service are assessed. The reliabilities and cost will provide a baseline for use in trade studies of future flight control system design.

Source record↗

Demonstration Advanced Avionics System (DAAS), Phase 1

Demonstration advanced anionics system (DAAS) function description, hardware description, operational evaluation, and failure mode and effects analysis (FMEA) are provided. Projected advanced avionics system (PAAS) description, reliability analysis, cost analysis, maintainability analysis, and modularity analysis are discussed.

Bailey, A. J.↗

The Exploration of Mars Launch and Assembly Simulation

Advancing human exploration of space beyond Low Earth Orbit, and ultimately to Mars, is of great interest to NASA, other organizations, and space exploration advocates. Various strategies for getting to Mars have been proposed. These include NASA's Design Reference Architecture 5.0, a near-term flyby of Mars advocated by the group Inspiration Mars, and potential options developed for NASA's Evolvable Mars Campaign. Regardless of which approach is used to get to Mars, they all share a need to visualize and analyze their proposed campaign and evaluate the feasibility of the launch and on-orbit assembly segment of the campaign. The launch and assembly segment starts with flight hardware manufacturing and ends with final departure of a Mars Transfer Vehicle (MTV), or set of MTVs, from an assembly orbit near Earth. This paper describes a discrete event simulation based strategic visualization and analysis tool that can be used to evaluate the launch campaign reliability of any proposed strategy for exploration beyond low Earth orbit. The input to the simulation can be any manifest of multiple launches and their associated transit operations between Earth and the exploration destinations, including Earth orbit, lunar orbit, asteroids, moons of Mars, and ultimately Mars. The simulation output includes expected launch dates and ascent outcomes i.e., success or failure. Running 1,000 replications of the simulation provides the capability to perform launch campaign reliability analysis to determine the probability that all launches occur in a timely manner to support departure opportunities and to deliver their payloads to the intended orbit. This allows for quantitative comparisons between alternative scenarios, as well as the capability to analyze options for improving launch campaign reliability. Results are presented for representative strategies.

Cates, Grant↗

The Exploration of Mars Launch and Assembly Simulation

Advancing human exploration of space beyond Low Earth Orbit, and ultimately to Mars, is of great interest to NASA, other organizations, and space exploration advocates. Various strategies for getting to Mars have been proposed. These include NASA's Design Reference Architecture 5.0, a near-term flyby of Mars advocated by the group Inspiration Mars, and potential options developed for NASA's Evolvable Mars Campaign. Regardless of which approach is used to get to Mars, they all share a need to visualize and analyze their proposed campaign and evaluate the feasibility of the launch and on-orbit assembly segment of the campaign. The launch and assembly segment starts with flight hardware manufacturing and ends with final departure of a Mars Transfer Vehicle (MTV), or set of MTVs, from an assembly orbit near Earth. This paper describes a discrete event simulation based strategic visualization and analysis tool that can be used to evaluate the launch campaign reliability of any proposed strategy for exploration beyond low Earth orbit. The input to the simulation can be any manifest of multiple launches and their associated transit operations between Earth and the exploration destinations, including Earth orbit, lunar orbit, asteroids, moons of Mars, and ultimately Mars. The simulation output includes expected launch dates and ascent outcomes i.e., success or failure. Running 1,000 replications of the simulation provides the capability to perform launch campaign reliability analysis to determine the probability that all launches occur in a timely manner to support departure opportunities and to deliver their payloads to the intended orbit. This allows for quantitative comparisons between alternative scenarios, as well as the capability to analyze options for improving launch campaign reliability. Results are presented for representative strategies.

Cates, Grant↗

Premium quality 5A1-2.5 Sn ELI titanium production

Preliminary design and reliability analysis conducted on the turbopump for the NERVA 75,000 full flow cycle engine, indicated that the turbopump bearings were the most critical turbopump parts in meeting the 10 hour life at the required turbopump reliability of .99978. The analysis revealed that significant reductions (approximately a factor of 3.25) in bearing loads would be achieved by fabricating the rotating parts from titanium in lieu of A286 or 718. This is basically due to the difference in density of the materials and the resulting mass effect on the location of the first and second stick mode critical speeds. For the selected rotor configuration, the lighter material has a first critical speed at approximately 36,000 rpm, while that of the heavier material has a first critical at approximately 27,000 rpm. As the operating range of the turbopump is from 0 to 30,000 rpm, the heavier material would have a stick mode critical in the operating range.

Dessau, P. P.↗