Search NASA⌕ Search

SEARCH · Search NASA

Results for “System reliability”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 775 records · Page 43

Reliability models for Space Station power system

This paper presents a methodology for the reliability evaluation of Space Station power system. The two options considered are the photovoltaic system and the solar dynamic system. Reliability models for both of these options are described along with the methodology for calculating the reliability indices.

Singh, C.↗

Reliability and maintainability assessment factors for reliable fault-tolerant systems

A long term goal of the NASA Langley Research Center is the development of a reliability assessment methodology of sufficient power to enable the credible comparison of the stochastic attributes of one ultrareliable system design against others. This methodology, developed over a 10 year period, is a combined analytic and simulative technique. An analytic component is the Computer Aided Reliability Estimation capability, third generation, or simply CARE III. A simulative component is the Gate Logic Software Simulator capability, or GLOSS. The numerous factors that potentially have a degrading effect on system reliability and the ways in which these factors that are peculiar to highly reliable fault tolerant systems are accounted for in credible reliability assessments. Also presented are the modeling difficulties that result from their inclusion and the ways in which CARE III and GLOSS mitigate the intractability of the heretofore unworkable mathematics.

Bavuso, S. J.↗

Reliability and coverage analysis of non-repairable fault-tolerant memory systems

A method was developed for the construction of probabilistic state-space models for nonrepairable systems. Models were developed for several systems which achieved reliability improvement by means of error-coding, modularized sparing, massive replication and other fault-tolerant techniques. From the models developed, sets of reliability and coverage equations for the systems were developed. Comparative analyses of the systems were performed using these equation sets. In addition, the effects of varying subunit reliabilities on system reliability and coverage were described. The results of these analyses indicated that a significant gain in system reliability may be achieved by use of combinations of modularized sparing, error coding, and software error control. For sufficiently reliable system subunits, this gain may far exceed the reliability gain achieved by use of massive replication techniques, yet result in a considerable saving in system cost.

Cox, G. W.↗

Composite, vacuum-jacketed tubing replaces bellows in cryogenic systems

For reliability control of high pressure cryogenic systems, one or more 90 degree elbow expansion devices are substituted for the metal bellows normally used. The device consists of a conducting tube inside a support tube, with the space between the tubes evacuated for insulation.

Calvert, H. F.↗

System life and reliability modeling for helicopter transmissions

A computer program which simulates life and reliability of helicopter transmissions is presented. The helicopter transmissions may be composed of spiral bevel gear units and planetary gear units - alone, in series or in parallel. The spiral bevel gear units may have either single or dual input pinions, which are identical. The planetary gear units may be stepped or unstepped and the number of planet gears carried by the planet arm may be varied. The reliability analysis used in the program is based on the Weibull distribution lives of the transmission components. The computer calculates the system lives and dynamic capacities of the transmission components and the transmission. The system life is defined as the life of the component or transmission at an output torque at which the probability of survival is 90 percent. The dynamic capacity of a component or transmission is defined as the output torque which can be applied for one million output shaft cycles for a probability of survival of 90 percent. A complete summary of the life and dynamic capacity results is produced by the program.

Savage, M.↗

A High-Reliability Photoelectric Detection System for Mars Sample Return’s Orbiting Sample

The Mars Sample Return campaign is an endeavor of unprecedented technological complexity and coordination that attempts to answer fundamental questions about the habitability of Mars by returning the first samples of Martian material to Earth for analysis. The third mission in the campaign consists of the NASA-provided Capture, Containment, and Return System (CCRS) onboard the European Space Agency’s Earth Return Orbiter, which will retrieve the Orbiting Sample (OS) container from its orbit around Mars. Retrieving a passive sample container from a planetary orbit has never been attempted by any spacecraft and requires the development of new technology to succeed in this ambitious task. This paper introduces the high-reliability Capture Sensor Suite (CSS), a novel optical detection system that provides CCRS with the capability to autonomously detect the OS as it is captured. This article will discuss the challenges and requirements for the fault-tolerant design of the CSS.

planetary sampling↗

Reliability and performance evaluation of systems containing embedded rule-based expert systems

A method for evaluating the reliability of real-time systems containing embedded rule-based expert systems is proposed and investigated. It is a three stage technique that addresses the impact of knowledge-base uncertainties on the performance of expert systems. In the first stage, a Markov reliability model of the system is developed which identifies the key performance parameters of the expert system. In the second stage, the evaluation method is used to determine the values of the expert system's key performance parameters. The performance parameters can be evaluated directly by using a probabilistic model of uncertainties in the knowledge-base or by using sensitivity analyses. In the third and final state, the performance parameters of the expert system are combined with performance parameters for other system components and subsystems to evaluate the reliability and performance of the complete system. The evaluation method is demonstrated in the context of a simple expert system used to supervise the performances of an FDI algorithm associated with an aircraft longitudinal flight-control system.

Beaton, Robert M.↗

Linking Plant and Microbial Traits to Soil Carbon for Reliable and Resilient Bioenergy Systems

Bioenergy systems in the United States offer a dual opportunity to supply renewable feedstocks while enhancing ecosystem services such as hydrologic regulation, erosion control, and soil carbon (C) storage. National assessments highlight the potential to grow perennial energy crops to improve soil function and ecosystem resilience. Realizing this potential requires understanding the ecological mechanisms that govern how C is added, transformed, and stabilized in soils. Plant traits determine the quantity, depth, and chemistry of organic inputs, while microbial processes—including carbon use efficiency, necromass formation, and trophic interactions—mediate their transformation and partitioning among soil carbon pools. These biological pathways are shaped by soil physical and chemical properties, including aggregation, texture, and mineralogy, and by environmental drivers such as temperature, moisture, and disturbance, leading to context-dependent outcomes across landscapes. Management practices that diversify feedstocks, minimize disturbance, and maintain soil cover can promote both biomass production and C retention, while microbial amendments and rhizosphere engineering offer emerging, but often context-dependent, tools to optimize plant–microbe interactions. Trade-offs between biomass yield and soil carbon storage may arise when systems favor rapid aboveground productivity at the expense of belowground inputs and microbial processing, underscoring the importance of trait combinations that support both functions. Advances in monitoring, reporting, and verification—spanning precision agriculture, remote sensing, and biosensing—are improving predictive capacity through microbial-explicit process models and model–experiment (ModEx) frameworks. By connecting soil, plant, and microbial processes with advances in modeling and biosensing, this review outlines research priorities focused on trait-based parameterization and ModEx integration. These priorities will support the design of bioenergy systems that are both reliable and resilient, enhancing renewable energy production and ecosystem sustainability.

bioenergy systems↗

Design and Analysis of a Flexible, Reliable Deep Space Life Support System

This report describes a flexible, reliable, deep space life support system design approach that uses either storage or recycling or both together. The design goal is to provide the needed life support performance with the required ultra reliability for the minimum Equivalent System Mass (ESM). Recycling life support systems used with multiple redundancy can have sufficient reliability for deep space missions but they usually do not save mass compared to mixed storage and recycling systems. The best deep space life support system design uses water recycling with sufficient water storage to prevent loss of crew if recycling fails. Since the amount of water needed for crew survival is a small part of the total water requirement, the required amount of stored water is significantly less than the total to be consumed. Water recycling with water, oxygen, and carbon dioxide removal material storage can achieve the high reliability of full storage systems with only half the mass of full storage and with less mass than the highly redundant recycling systems needed to achieve acceptable reliability. Improved recycling systems with lower mass and higher reliability could perform better than systems using storage.

reliability↗

The implementation and use of Ada on distributed systems with high reliability requirements

The use and implementation of Ada (a trade mark of the US Dept. of Defense) in distributed environments in which the hardware are assumed to be unreliable were investigated. The possibility that a distributed system is programmed entirely in Ada so that the individual tasks of the system are unconcerned with which processors they are executing on and failures occurring in the underlying hardware were examined.

Knight, J. C.↗

ROBUS-2: A Fault-Tolerant Broadcast Communication System

The Reliable Optical Bus (ROBUS) is the core communication system of the Scalable Processor-Independent Design for Enhanced Reliability (SPIDER), a general-purpose fault-tolerant integrated modular architecture currently under development at NASA Langley Research Center. The ROBUS is a time-division multiple access (TDMA) broadcast communication system with medium access control by means of time-indexed communication schedule. ROBUS-2 is a developmental version of the ROBUS providing guaranteed fault-tolerant services to the attached processing elements (PEs), in the presence of a bounded number of faults. These services include message broadcast (Byzantine Agreement), dynamic communication schedule update, clock synchronization, and distributed diagnosis (group membership). The ROBUS also features fault-tolerant startup and restart capabilities. ROBUS-2 is tolerant to internal as well as PE faults, and incorporates a dynamic self-reconfiguration capability driven by the internal diagnostic system. This version of the ROBUS is intended for laboratory experimentation and demonstrations of the capability to reintegrate failed nodes, dynamically update the communication schedule, and tolerate and recover from correlated transient faults.

Torres-Pomales, Wilfredo↗

From Data to Knowledge: A Graph-Based Reliability Approach to Assess System Health

With the goal of maximizing plant reliability and availability, complex systems such as nuclear power plants continuously monitor and record the performance and the health status of many components, assets, and systems. Such data may take the form of online monitoring data, condition reports, and maintenance reports and it carries the potential to provide system engineers with insights into anomalous behaviors or degradation trends as well as the possible causes behind them and to predict their direct consequences. The analysis of such data poses however few challenges. While some of these challenges are technical in nature (i.e., data are often distributed over several physical servers or databases), others are conceptual in nature (i.e., data elements come in different formats, numeric or textual), and measured values have different scales (e.g., vibration spectra and oil temperature). This paper directly tackles these challenges, and it focuses on the integration of all these data elements in order to assist plant system engineers in analyzing component, assets, and systems performances and optimize maintenance activities. This is performed by 1) extracting knowledge from textual data via technical language processing methods, and 2) quantifying system, asset, and component health from numeric condition-based data. We rely on model-based system engineering (MBSE) models of systems and assets to identify their architecture and functional (i.e., cause and effect) relations. Numeric and textual data elements are then associated with an MBSE graph element, based on their nature. This bonding of MBSE models and data elements constitutes a first-of-its-kind knowledge graph of a nuclear power plants system, with data elements being organized in a structured manner that enables system engineers to identify cause-effect trends in data elements and carry out appropriate actions in response.

97 MATHEMATICS AND COMPUTING↗

Models of failure in linear systems

The concept of reliability in systems theory is discussed. Emphasis is placed on identifying failed or degraded systems and constructing a strategy to control the system in the failed mode. Systems are assumed to be linear and transformations into the failed mode occur smoothly. The family into which a nominal system is embedded is examined along with the relations within the family. The class of systems into which a given system can be algebraically transformed is investigated.

Martin, C. F.↗

Reliability of Fault Tolerant Control Systems

This paper reports Part II of a two part effort that is intended to delineate the relationship between reliability and fault tolerant control in a quantitative manner. Reliability properties peculiar to fault-tolerant control systems are emphasized, such as the presence of analytic redundancy in high proportion, the dependence of failures on control performance, and high risks associated with decisions in redundancy management due to multiple sources of uncertainties and sometimes large processing requirements. As a consequence, coverage of failures through redundancy management can be severely limited. The paper proposes to formulate the fault tolerant control problem as an optimization problem that maximizes coverage of failures through redundancy management. Coverage modeling is attempted in a way that captures its dependence on the control performance and on the diagnostic resolution. Under the proposed redundancy management policy, it is shown that an enhanced overall system reliability can be achieved with a control law of a superior robustness, with an estimator of a higher resolution, and with a control performance requirement of a lesser stringency.

Wu, N. Eva↗

Evaluation Applied to Reliability Analysis of Reconfigurable, Highly Reliable, Fault-Tolerant, Computing Systems for Avionics

Emulation techniques are proposed as a solution to a difficulty arising in the analysis of the reliability of highly reliable computer systems for future commercial aircraft. The difficulty, viz., the lack of credible precision in reliability estimates obtained by analytical modeling techniques are established. The difficulty is shown to be an unavoidable consequence of: (1) a high reliability requirement so demanding as to make system evaluation by use testing infeasible, (2) a complex system design technique, fault tolerance, (3) system reliability dominated by errors due to flaws in the system definition, and (4) elaborate analytical modeling techniques whose precision outputs are quite sensitive to errors of approximation in their input data. The technique of emulation is described, indicating how its input is a simple description of the logical structure of a system and its output is the consequent behavior. The use of emulation techniques is discussed for pseudo-testing systems to evaluate bounds on the parameter values needed for the analytical techniques.

Migneault, G. E.↗

Radiation Single Event Effects (SEE) Impact on Complex Avionics Architecture Reliability

The NASA Engineering and Safety Center (NESC) has an urgent need to understand how system-level reliability of an avionics architecture is compromised when portions of the architecture are temporarily unavailable due to single event effects (SEE). The proposed activity parametrically evaluated these SEE impacts on system reliability based on mission duration, upset rate and recovery times for a representative redundant architecture. The key stakeholders for this study are NASA programs and projects that expect to use avionics architectures with electrical, electronic and electromechanical (EEE) parts susceptible to SEE when exposed to the mission expected radiation environment.

Hodson, Robert F.↗