Search NASA⌕ Search

SEARCH · Search NASA

Results for “Failure Rate”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

A nonparametric software-reliability growth model

The authors (1985) previously introduced a nonparametric model for software-reliability growth which is based on complete monotonicity of the failure rate. The authors extend the completely monotone software model by developing a method for providing long-range predictions of reliability growth, based on the model. They derive upper and lower bounds on extrapolation of the failure rate and the mean function. These are then used to obtain estimates for the future software failure rate and the mean future number of failures. Preliminary evaluation indicates that the method is competitive with parametric approaches, while being more robust.

Sofer, Ariela↗

Surrogate oracles, generalized dependency and simpler models

Software reliability models require the sequence of interfailure times from the debugging process as input. It was previously illustrated that using data from replicated debugging could greatly improve reliability predictions. However, inexpensive replication of the debugging process requires the existence of a cheap, fast error detector. Laboratory experiments can be designed around a gold version which is used as an oracle or around an n-version error detector. Unfortunately, software developers can not be expected to have an oracle or to bear the expense of n-versions. A generic technique is being investigated for approximating replicated data by using the partially debugged software as a difference detector. It is believed that the failure rate of each fault has significant dependence on the presence or absence of other faults. Thus, in order to discuss a failure rate for a known fault, the presence or absence of each of the other known faults needs to be specified. Also, in simpler models which use shorter input sequences without sacrificing accuracy are of interest. In fact, a possible gain in performance is conjectured. To investigate these propositions, NASA computers running LIC (RTI) versions are used to generate data. This data will be used to label the debugging graph associated with each version. These labeled graphs will be used to test the utility of a surrogate oracle, to analyze the dependent nature of fault failure rates and to explore the feasibility of reliability models which use the data of only the most recent failures.

Wilson, Larry↗

A new algorithm for finding survival coefficients employed in reliability equations

Product reliabilities are predicted from past failure rates and reasonable estimate of future failure rates. Algorithm is used to calculate probability that product will function correctly. Algorithm sums the probabilities of each survival pattern and number of permutations for that pattern, over all possible ways in which product can survive.

Bouricius, W. G.↗

Comparative analysis of different configurations of PLC-based safety systems from reliability point of view

The study of a comparative analysis of distinct multiplex and fault-tolerant configurations for a PLC-based safety system from a reliability point of view is presented. It considers simplex, duplex and fault-tolerant triple redundancy configurations. The standby unit in case of a duplex configuration has a failure rate which is k times the failure rate of the standby unit, the value of k varying from 0 to 1. For distinct values of MTTR and MTTF of the main unit, MTBF and availability for these configurations are calculated. The effect of duplexing only the PLC module or only the sensors and the actuators module, on the MTBF of the configuration, is also presented. The results are summarized and merits and demerits of various configurations under distinct environments are discussed.

Tapia, Moiez A.↗

Scaled CMOS Technology Reliability Users Guide

The desire to assess the reliability of emerging scaled microelectronics technologies through faster reliability trials and more accurate acceleration models is the precursor for further research and experimentation in this relevant field. The effect of semiconductor scaling on microelectronics product reliability is an important aspect to the high reliability application user. From the perspective of a customer or user, who in many cases must deal with very limited, if any, manufacturer's reliability data to assess the product for a highly-reliable application, product-level testing is critical in the characterization and reliability assessment of advanced nanometer semiconductor scaling effects on microelectronics reliability. A methodology on how to accomplish this and techniques for deriving the expected product-level reliability on commercial memory products are provided.Competing mechanism theory and the multiple failure mechanism model are applied to the experimental results of scaled SDRAM products. Accelerated stress testing at multiple conditions is applied at the product level of several scaled memory products to assess the performance degradation and product reliability. Acceleration models are derived for each case. For several scaled SDRAM products, retention time degradation is studied and two distinct soft error populations are observed with each technology generation: early breakdown, characterized by randomly distributed weak bits with Weibull slope (beta)=1, and a main population breakdown with an increasing failure rate. Retention time soft error rates are calculated and a multiple failure mechanism acceleration model with parameters is derived for each technology. Defect densities are calculated and reflect a decreasing trend in the percentage of random defective bits for each successive product generation. A normalized soft error failure rate of the memory data retention time in FIT/Gb and FIT/cm2 for several scaled SDRAM generations is presented revealing a power relationship. General models describing the soft error rates across scaled product generations are presented. The analysis methodology may be applied to other scaled microelectronic products and their key parameters.

Microelectronics Reliability↗

Four Problematic Methods in Reliability Analysis

Some basic methods used in reliability analysis are problematic because they produce incorrect and overoptimistic predictions. Initially gratifying forecasts are often invalidated by testing and operational experience. The problematic methods in reliability analysis include estimating the system failure rate as the sum of component failure rates, assuming that reliability growth continues indefinitely during testing, overestimating the benefits of redundancy, and using the fault tolerance count instead of a detailed reliability analysis. Reliability analysis can produce more optimism than accuracy. This bug may now be a feature. The optimistic bias inevitable in project planning should be corrected by realistic reliability analysis that reflects relevant experience. That the repeated poor performance of reliability analysis is found to be surprising suggests willful blindness. Rigorous methods and impartial critical review are necessary to improve reliability analysis.

Reliability analysis↗

Four Problematic Methods in Reliability Analysis

Some basic methods used in reliability analysis are problematic because they produce incorrect and overoptimistic predictions. Initially gratifying forecasts are often invalidated by testing and operational experience. The problematic methods in reliability analysis include estimating the system failure rate as the sum of component failure rates, assuming that reliability growth continues indefinitely during testing, overestimating the benefits of redundancy, and using the fault tolerance count instead of a detailed reliability analysis. Reliability analysis can produce more optimism than accuracy. This bug may now be a feature. The optimistic bias inevitable in project planning should be corrected by realistic reliability analysis that reflects relevant experience. That the repeated poor performance of reliability analysis is found to be surprising suggests willful blindness. Rigorous methods and impartial critical review are necessary to improve reliability analysis.

Reliability analysis↗

Remote operation of an orbital maneuvering vehicle in simulated docking maneuvers

Simulated docking maneuvers were performed to assess the effect of initial velocity on docking failure rate, mission duration, and delta v (fuel consumption). Subjects performed simulated docking maneuvers of an orbital maneuvering vehicle (OMV) to a space station. The effect of the removal of the range and rate displays (simulating a ranging instrumentation failure) was also examined. Naive subjects were capable of achieving a high success rate in performing simulated docking maneuvers without extensive training. Failure rate was a function of individual differences; there was no treatment effect on failure rate. The amount of time subjects reserved for final approach increased with starting velocity. Piloting of docking maneuvers was not significantly affected in any way by the removal of range and rate displays. Radial impulse was significant both by subject and by treatment. NASA's 0.1 percent rule, dictating an approach rate no greater than 0.1 percent of the range, is seen to be overly conservative for nominal docking missions.

Brody, Adam R.↗

The effects of inherent flaws on the time and rate dependent failure of adhesively bonded joints

Inherent flaws, as well as the effects of rate and time, are shown by tests on viscoelastic adhesive-bonded single lap joints to be as critical in joint failure as environmental and stress concentration effects, with random inherent flaws and loading rate changes resulting in an up to 40% reduction in joint strength. It is also found that the asymptotic creep stress, below which no delayed failure may occur, may under creep loading be as much as 45% less than maximum adhesive strength. Attention is given to test results for the case of titanium-LARC-3 adhesive single-lap specimens.

Sancaktar, E.↗

Trends in reliability modeling technology for fault tolerant systems

Developments in reliability modeling for large fault tolerant avionic computing systems are presented. Issues of state size and complexity, fault coverage, and practical computation are addressed. A two-fold developmental effort is described based on the structural and fault coverage modeling approaches. A technique which was successfully applied to an 865 state pure death stationary Markov model is presented. Of particular interest is a short computer program which executes very quickly to produce reliability results of a large state space model. This model also incorporates fault coverage states for processor, memory, and bus line replaceable units. A second structural reliability modeling scheme is aimed at solving nonstationary Markov models. This technique provides the tool required for studying the reliability of systems with nonconstant failure rates and includes intermittent/transient faults, electronic hardware which exhibits decreasing failure rates, and hydromechanical devices which typically have wearout failure mechanisms. Several aspects of fault coverage, including modeling and data measurement of intermittent/transient faults and latent faults, are elucidated and illustrated. The CARE II (computer-aided reliability estimation) coverage is presented and shortcomings to be eliminated are discussed.

Bavuso, S. J.↗

Solar Photovoltaic (PV) Damage Assessment After Typhoon Mawar: Findings and Recommendations for Resilient PV on Guam

A team from the National Renewable Energy Laboratory (NREL) visited Guam in August 2023 to assess failure modes of solar photovoltaic (PV) systems after Typhoon Mawar and to provide recommendations to increase the resilience of PV systems on Guam. The team visited 30 systems: commercial and utility scale, and rooftop and ground-mounted. The team observed systems with no apparent damage, as well as systems that were completely lost. Systems fared very well overall. The average failure rate of rooftop systems was 18%, with a median failure rate of 2%, meaning the few systems that suffered total loss pulled up the average. Only eight 8 of the 25 rooftop systems suffered more than 5% damage. All ground-mounted systems suffered less than 0.5% damage, aside from a carport that lost 16% of its modules. PV systems at Andersen Air Force Base suffered 5% damage on average, with a median system failure of 0.6%. In almost all cases, failures were the result of: (1) Inadequate clamping of the module frame to the mount, (2) Module mounting clamps rotating out of underlying support rail (i.e., T-bolt that rotates free at less than 60 degrees of rotation), (3) An object hitting the panel resulting in a fracture, and in some cases leading to a cascading failure of several more panels, and (4) Excessive tilt angle (in Guam, greater than 5 degrees can be a risk due to wind speed, and power production trade-offs are insignificant).

14 SOLAR ENERGY↗

Test effectiveness study report: An analytical study of system test effectiveness and reliability growth of three commercial spacecraft programs

Failure data from 16 commercial spacecraft were analyzed to evaluate failure trends, reliability growth, and effectiveness of tests. It was shown that the test programs were highly effective in ensuring a high level of in-orbit reliability. There was only a single catastrophic problem in 44 years of in-orbit operation on 12 spacecraft. The results also indicate that in-orbit failure rates are highly correlated with unit and systems test failure rates. The data suggest that test effectiveness estimates can be used to guide the content of a test program to ensure that in-orbit reliability goals are achieved.

Feldstein, J. F.↗

Assuring reliability program effectiveness.

An attempt is made to provide simple identification and description of techniques that have proved to be most useful either in developing a new product or in improving reliability of an established product. The first reliability task is obtaining and organizing parts failure rate data. Other tasks are parts screening, tabulation of general failure rates, preventive maintenance, prediction of new product reliability, and statistical demonstration of achieved reliability. Five principal tasks for improving reliability involve the physics of failure research, derating of internal stresses, control of external stresses, functional redundancy, and failure effects control. A final task is the training and motivation of reliability specialist engineers.

Ball, L. W.↗

Developing Ultra Reliable Life Support for the Moon and Mars

Recycling life support systems can achieve ultra reliability by using spares to replace failed components. The added mass for spares is approximately equal to the original system mass, provided the original system reliability is not very low. Acceptable reliability can be achieved for the space shuttle and space station by preventive maintenance and by replacing failed units, However, this maintenance and repair depends on a logistics supply chain that provides the needed spares. The Mars mission must take all the needed spares at launch. The Mars mission also must achieve ultra reliability, a very low failure rate per hour, since it requires years rather than weeks and cannot be cut short if a failure occurs. Also, the Mars mission has a much higher mass launch cost per kilogram than shuttle or station. Achieving ultra reliable space life support with acceptable mass will require a well-planned and extensive development effort. Analysis must define the reliability requirement and allocate it to subsystems and components. Technologies, components, and materials must be designed and selected for high reliability. Extensive testing is needed to ascertain very low failure rates. Systems design should segregate the failure causes in the smallest, most easily replaceable parts. The systems must be designed, produced, integrated, and tested without impairing system reliability. Maintenance and failed unit replacement should not introduce any additional probability of failure. The overall system must be tested sufficiently to identify any design errors. A program to develop ultra reliable space life support systems with acceptable mass must start soon if it is to produce timely results for the moon and Mars.

Jones, Harry W.↗

Module Hipot and ground continuity test results

Hipot (high voltage potential) and module frame continuity tests of solar energy conversion modules intended for deployment into large arrays are discussed. The purpose of the tests is to reveal potentially hazardous voltage conditions in installed modules, and leakage currents that may result in loss of power or cause ground fault system problems, i.e., current leakage potential and leakage voltage distribution. The tests show a combined failure rate of 36% (69% when environmental testing is included). These failure rates are believed easily corrected by greater care in fabrication.

Griffith, J. S.↗

The effect of initial velocity on manually controlled remote docking of an orbital maneuvering vehicle (OMV) to a space station

Simulated docking maneuvers were performed to assess the effect of initial velocity on docking failure rate, mission duration, and total impulse (fuel consumption). The effect of the removal of the range and rate displays was also examined. Since duration and impulse decrease and increase respectively with increases in initial velocity, two parameters were created by subtracting a reference value from each. These values were termed 'reserve time' and 'radial impulse'. Naive subjects were capable of achieving a high success rate in performing simulated docking maneuvers without extensive experience, and failure rate did not significantly increase with increased velocity. The amount of time pilots reserved for final approach increased with starting velocity. Piloting of docking maneuvers was not significantly affected in any way by the removal of range and rate displays. Values for reserve time, and radial impulse were lowest for docking maneuvers begun at the lowest initial velocity.

Brody, Adam R.↗

Ultra Reliable Closed Loop Life Support for Long Space Missions

Spacecraft human life support systems can achieve ultra reliability by providing sufficient spares to replace all failed components. The additional mass of spares for ultra reliability is approximately equal to the original system mass, provided that the original system reliability is not too low. Acceptable reliability can be achieved for the Space Shuttle and Space Station by preventive maintenance and by replacing failed units. However, on-demand maintenance and repair requires a logistics supply chain in place to provide the needed spares. In contrast, a Mars or other long space mission must take along all the needed spares, since resupply is not possible. Long missions must achieve ultra reliability, a very low failure rate per hour, since they will take years rather than weeks and cannot be cut short if a failure occurs. Also, distant missions have a much higher mass launch cost per kilogram than near-Earth missions. Achieving ultra reliable spacecraft life support systems with acceptable mass will require a well-planned and extensive development effort. Analysis must determine the reliability requirement and allocate it to subsystems and components. Ultra reliability requires reducing the intrinsic failure causes, providing spares to replace failed components and having "graceful" failure modes. Technologies, components, and materials must be selected and designed for high reliability. Long duration testing is needed to confirm very low failure rates. Systems design should segregate the failure causes in the smallest, most easily replaceable parts. The system must be designed, developed, integrated, and tested with system reliability in mind. Maintenance and reparability of failed units must not add to the probability of failure. The overall system must be tested sufficiently to identify any design errors. A program to develop ultra reliable space life support systems with acceptable mass should start soon since it must be a long term effort.

Jones, Harry W.↗