Search NASA⌕ Search

SEARCH · Search NASA

Results for “SYSTEM FAILURE”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 235 records · Page 13

Fault Diagnosis of Power Components with Reliability Assessment in Extraterrestrial Microgrids

This research investigates the possible failures caused by aging and other environmental and external factors that could significantly impact the performance of extraterrestrial power systems. Additionally, it presents a reliability assessment model for the space microgrid based on fault tree analysis (FTA). The reliability assessment model developed in this paper represents a tool that can be used by engineers to harden the system design for operational and economic benefits. To improve the reliability of the system, this work provides a broad review of the different fault detection and diagnosis (FDD) algorithms used for power microgrids and space applications. Using data sets from the Habitat Simulator developed through the NASA-funded Resilient Extraterrestrial Habitat Institute, this paper compares the applicability and accuracy of the different FDD methods. The primary FDD approach proposed and assessed in this work is based on the Markov reliability model. It predicts and detects future faults in the space microgrids by using past data samples and categorizing them into different classes. Data-driven-based models such as artificial neural networks are also investigated, tested, and evaluated using simulation data sets. According to the simulation results and the broad FDD algorithm comparison, this study provides the crew or maintenance engineers with a clear methodology to detect and localize power system failures.

Leila Chebbo↗

Preventing Spacecraft Failures Due to Tribological Problems

Many mechanical failures that occur on spacecraft are caused by tribological problems. This publication presents a study that was conducted by the author on various preventatives, analyses, controls and tests (PACTs) that could be used to prevent spacecraft mechanical system failure. A matrix is presented in the paper that plots tribology failure modes versus various PACTs that should be performed before a spacecraft is launched in order to insure success. A strawman matrix was constructed by the author and then was sent out to industry and government spacecraft designers, scientists and builders of spacecraft for their input. The final matrix is the result of their input. In addition to the matrix, this publication describes the various PACTs that can be performed and some fundamental knowledge on the correct usage of lubricants for spacecraft applications. Even though the work was done specifically to prevent spacecraft failures the basic methodology can be applied to other mechanical system areas.

Fusaro, Robert L.↗

Synthetic Failure Mode Generation for Resilience Analysis and Failure Mechanism Discovery

Traditional risk-based design processes seek to mitigate operational hazards by manually identifying possible faults and corresponding mitigation strategies—a tedious process which critically relies on the designer’s limited knowledge. Resilience-based design, on the other hand, seeks to embody generic hazard-mitigating properties in the system to mitigate unknown hazards, often by modelling the system's response to potential hazardous events. This work adapts this approach to the traditional risk-based design process to synthetically generate hazardous modes, by representing them as a unique combination of internal component health-states which can then be injected and simulated in a model of the system failure dynamics. The design process may then reduce the risk of unknown internal hazards by iteratively mitigating the effects of these modes. The performance of this approach is evaluated in a model of an autonomous rover, where cluster analysis shows that elaborating the space of synthetic faults in the drive system using this approach uncovers a wider range of possible hazardous trajectories and failure consequences within each trajectory. However, this increase in hazard information comes at a high computational expense, highlighting the need for advanced, efficient methods to search and sample the hazard space.

Simulation↗

Synthetic Failure Mode Generation for Resilience Analysis and Failure Mechanism Discovery

Traditional risk-based design processes seek to mitigate operational hazards by manually identifying possible faults and corresponding mitigation strategies—a tedious process which critically relies on the designer’s limited knowledge. Resilience-based design, on the other hand, seeks to embody generic hazard-mitigating properties in the system to mitigate unknown hazards, often by modelling the system's response to potential hazardous events. This work adapts this approach to the traditional risk-based design process to synthetically generate hazardous modes, by representing them as a unique combination of internal component health-states which can then be injected and simulated in a model of the system failure dynamics. The design process may then reduce the risk of unknown internal hazards by iteratively mitigating the effects of these modes. The performance of this approach is evaluated in a model of an autonomous rover, where cluster analysis shows that elaborating the space of synthetic faults in the drive system using this approach uncovers a wider range of possible hazardous trajectories and failure consequences within each trajectory. However, this increase in hazard information comes at a high computational expense, highlighting the need for advanced, efficient methods to search and sample the hazard space.

Simulation↗

Quantifying Pilot Contribution to Flight Safety During an In-Flight Airspeed Failure

Accident statistics cite the flight crew as a causal factor in over 60% of large transport fatal accidents. Yet a well-trained and well-qualified crew is acknowledged as the critical center point of aircraft systems safety and an integral component of the entire commercial aviation system. A human-in-the-loop test was conducted using a Level D certified Boeing 737-800 simulator to evaluate the pilot's contribution to safety-of-flight during routine air carrier flight operations and in response to system failures. To quantify the human's contribution, crew complement was used as an independent variable in a between-subjects design. This paper details the crew's actions and responses while dealing with an in-flight airspeed failure. Accident statistics often cite flight crew error (Baker, 2001) as the primary contributor in accidents and incidents in transport category aircraft. However, the Air Line Pilots Association (2011) suggests "a well-trained and well-qualified pilot is acknowledged as the critical center point of the aircraft systems safety and an integral safety component of the entire commercial aviation system." This is generally acknowledged but cannot be verified because little or no quantitative data exists on how or how many accidents/incidents are averted by crew actions. Anecdotal evidence suggest crews handle failures on a daily basis and Aviation Safety Action Program data generally supports this assertion, even if the data is not released to the public. However without hard evidence, the contribution and means by which pilots achieve safety of flight is difficult to define. Thus, ways to improve the human ability to contribute or overcome deficiencies are ill-defined.

Etherington, Timothy J.↗

Motion simulator study of longitudinal stability requirements for large delta wing transport airplanes during approach and landing with stability augmentation systems failed

A ground-based simulator investigation was conducted in preparation for and correlation with an-flight simulator program. The objective of these studies was to define minimum acceptable levels of static longitudinal stability for landing approach following stability augmentation systems failures. The airworthiness authorities are presently attempting to establish the requirements for civil transports with only the backup flight control system operating. Using a baseline configuration representative of a large delta wing transport, 20 different configurations, many representing negative static margins, were assessed by three research test pilots in 33 hours of piloted operation. Verification of the baseline model to be used in the TIFS experiment was provided by computed and piloted comparisons with a well-validated reference airplane simulation. Pilot comments and ratings are included, as well as preliminary tracking performance and workload data.

Snyder, C. T.↗

Hybrid Exploration Agent Platform and Sensor Web System

A sensor web to collect the scientific data needed to further exploration is a major and efficient asset to any exploration effort. This is true not only for lunar and planetary environments, but also for interplanetary and liquid environments. Such a system would also have myriad direct commercial spin-off applications. The Hybrid Exploration Agent Platform and Sensor Web or HEAP-SW like the ANTS concept is a Sensor Web concept. The HEAP-SW is conceptually and practically a very different system. HEAP-SW is applicable to any environment and a huge range of exploration tasks. It is a very robust, low cost, high return, solution to a complex problem. All of the technology for initial development and implementation is currently available. The HEAP Sensor Web or HEAP-SW consists of three major parts, The Hybrid Exploration Agent Platforms or HEAP, the Sensor Web or SW and the immobile Data collection and Uplink units or DU. The HEAP-SW as a whole will refer to any group of mobile agents or robots where each robot is a mobile data collection unit that spends most of its time acting in concert with all other robots, DUs in the web, and the HEAP-SWs overall Command and Control (CC) system. Each DU and robot is, however, capable of acting independently. The three parts of the HEAP-SW system are discussed in this paper. The Goals of the HEAP-SW system are: 1) To maximize the amount of exploration enhancing science data collected; 2) To minimize data loss due to system malfunctions; 3) To minimize or, possibly, eliminate the risk of total system failure; 4) To minimize the size, weight, and power requirements of each HEAP robot; 5) To minimize HEAP-SW system costs. The rest of this paper discusses how these goals are attained.

Stoffel, A. William↗

Monitoring Distributed Real-Time Systems: A Survey and Future Directions

Runtime monitors have been proposed as a means to increase the reliability of safety-critical systems. In particular, this report addresses runtime monitors for distributed hard real-time systems. This class of systems has had little attention from the monitoring community. The need for monitors is shown by discussing examples of avionic systems failure. We survey related work in the field of runtime monitoring. Several potential monitoring architectures for distributed real-time systems are presented along with a discussion of how they might be used to monitor properties of interest.

Goodloe, Alwyn E.↗

Assessment of Crew Time for Maintenance and Repair Activities for Lunar Surface Missions

NASA is currently evaluating different methods to predict how much time crewmembers will spend conducting repair and maintenance activities on future space missions. As mission scope and spacecraft architectures change, understanding how crew repair and maintenance timelines are impacted by mission operations and technology changes is vital for future mission planning. Past work has been done using historical International Space Station (ISS) data to accurately predict crew habitation and operation timelines, resulting in the development of NASA’s Exploration Crew Time Model (ECTM). However, understanding crew maintenance and repair requirements has posed a unique challenge due to the complexity of available datasets, the probabilistic nature of sub-system failures, and the impacts of reliability growth on failure rates. This paper presents a methodology to collect and condition empirical repair and maintenance time data from available datasets, to extrapolate from that data to estimate projected maintenance and repair times for a lunar Surface Habitat (SH), and to assess how uncertainty in repair time could impact utilization time on the lunar surface. NASA ISS maintenance and crew time data are logged into two central databases: the Maintenance Data Collection (MDC) and the Operations Planning Timeline Integration System (OPTimIS). Separately, each of these two datasets capture only portions of the complete set of data required to generate an accurate assessment of crew time spent on maintenance activities at a sub-system level. To create a more useful crew time estimate for maintenance timelines, the authors developed a methodology to capture relevant data from each set and combine and utilize that data by linking crew time requirements to specific components. The authors compare the failure logs in the MDC to crew activity logs pulled from OPTimIS and then process the data to estimate required repair time for each failure and repair event. The entire maintenance activity dataset is then categorized based on the class of failed component to ensure a significant sample size for each class and accurate crew time estimates for any components lacking relevant data. This resultant component repair time data can be used in the future to generate Mean Time to Repair (MTTR) estimates and confidence intervals for each class of component based on a probabilistic distribution of documented maintenance events. These improved MTTR values can then be applied to candidate element sub-system architectures, along with component Mean Time Between Failure (MTBF) data to generate distributions for potential required system crew repair time estimates for a given mission. The authors applied these modeling methods to a case study of a crewed mission to the planned SH and produced expected corrective maintenance crew time distributions. The results produced an expected corrective maintenance crew time at over 24 hours per mission, and a maintenance crew time distribution that reflects the importance of planning for sufficient maintenance requirements each mission. Repair time distributions can then be used to develop more accurate crew schedules and to assess potential available utilization time.

Crew Time↗

Ferrographic and spectrographic analysis of oil sampled before and after failure of a jet engine

An experimental gas turbine engine was destroyed as a result of the combustion of its titanium components. Several engine oil samples (before and after the failure) were analyzed with a Ferrograph as well as plasma, atomic absorption, and emission spectrometers. The analyses indicated that a lubrication system failure was not a causative factor in the engine failure. Neither an abnormal wear mechanism, nor a high level of wear debris was detected in the oil sample from the engine just prior to the test in which the failure occurred. However, low concentrations of titanium were evident in this sample and samples taken earlier. After the failure, higher titanium concentrations were detected in oil samples taken from different engine locations. Ferrographic analysis indicated that most of the titanium was contained in spherical metallic debris after the failure.

Jones, W. R., Jr.↗

Flight control system development and flight test experience with the F-111 mission adaptive wing aircraft

The wing on the NASA F-111 transonic aircraft technology airplane was modified to provide flexible leading and trailing edge flaps. This wing is known as the mission adaptive wing (MAW) because aerodynamic efficiency can be maintained at all speeds. Unlike a conventional wing, the MAW has no spoilers, external flap hinges, or fairings to break the smooth contour. The leading edge flaps and three-segment trailing edge flaps are controlled by a redundant fly-by-wire control system that features a dual digital primary system architecture providing roll and symmetric commands to the MAW control surfaces. A segregated analog backup system is provided in the event of a primary system failure. This paper discusses the design, development, testing, qualification, and flight test experience of the MAW primary and backup flight control systems.

Larson, R. R.↗

Flight control system development and flight test experience with the F-111 mission adaptive wing aircraft

The wing on the NASA F-111 transonic aircraft technology airplane was modified to provide flexible leading and trailing edge flaps. This wing is known as the mission adaptive wing (MAW) because aerodynamic efficiency can be maintained at all speeds. Unlike a conventional wing, the MAW has no spoilers, external flap hinges, or fairings to break the smooth contour. The leading edge flaps and three-segment trailing edge flaps are controlled by a redundant fly-by-wire control system that features a dual digital primary system architecture providing roll and symmetric commands to the MAW control surfaces. A segregated analog backup system is provided in the event of a primary system failure. This paper discusses the design, development, testing, qualification, and flight test experience of the MAW primary and backup flight control systems.

Larson, R. R.↗

Nuclear Safety [Vol. 16, No. 2, March-April 1975]

Nuclear Safety covers significant developments in the field of nuclear safety. The scope is limited to topics relevant to the analysis and control of hazards associated with nuclear energy, operations involving fissionable materials, and the products of nuclear fission and their effects on the environment. Primary emphasis is on safety in reactor design, construction, and operation; however, safety considerations in regard to the entire fuel cycle, including fuel fabrication, spent-fuel processing, nuclear waste disposal, handling of radioisotopes, and environmental effects of these operations, are also treated. Table of Contents for this issue follows. Table of Contents for this issue follows. GENERAL SAFETY CONSIDERATIONS: 127 Quality Assurance in the Construction of Nuclear Power Plants by Sidney A. Bernsen, 141 1974 ANS Topical Meeting on Fast Reactor Safety by M. H. Fontana; CONTROL AND INSTRUMENTATION: 150 GBR-4 Protection Systems: Failures and Their Consequences by Peter Burgsmüller, J. J. Dekais, Albert Krähe, Raffaello Pignatelli, and Gottfried Vieider, 162 Standby Emergency Power Systems. Part 2—The Later Plants by E. W. Hagen; PLANT SAFETY FEATURES: 180 Radiotoxic Hazard Measure for Buried Solid Radioactive Waste by J. Hamstra, 190 The Thirteenth AEC Air-Cleaning Conference by D. W. Moeller, D. W. Underhill, and M. W. First, 203 Book Review: Nuclear Criticality Safety; CONSEQUENCES OF EFFLUENT RELEASE: 204 Environmental Radiation Effects of Nuclear Facilities in New York State by M. S. Terpilak and B. L. Jorgensen, 222 Book Review: Thermal Ecology; OPERATING EXPERIENCES: 223 Set-Point Drift in Nuclear Power-Plant Safety-Related Instrumentation Adapted by the Nuclear Safety Staff, 224 Diesel-Generator Operating Experience at Nuclear Power Plants, 227 Summary of Operating U. S. Power Reactors as of Jan. 1, 1975, 232 Selected Safety-Related Occurrences Reported in November and December 1974 Compiled by William R. Casto, 235 Recent Occurrences at Nuclear Reactors and Their Causes Compiled by William R. Casto; CURRENT EVENTS: 243 General Administrative Activities Compiled by Wm. B. Cottrell, 251 Action on Power-Reactor Projects Undergoing Regulatory Review or Consideration Compiled by Wm. B. Cottrell, 266 Action on Nonreactor Projects Undergoing Regulatory Review or Consideration Compiled by Wm. B Cottrell, 268 Proposed Rule Changes as of Jan. 1, 1975; MISCELLANY: 149 Course in Italy on High-Energy Radiation Dosimetry and Protection (Announcement), 250 Course at Northwestern on Safety of Light-Water-Cooled Nuclear Power Plants (Announcement), 266 Symposium of the Combined Effects on the Environment of Radioactive, Chemical, and Thermal Releases from the Nuclear Industry (Announcement), 271 Short Course on Engineering for Extreme Winds and Tornadoes (Announcement), 272 Three 1-Week Courses at MIT on Nuclear Power-Reactor Safety (Announcement), 272 Harvard University Short Courses (Announcement).

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Phase plane displays detect incipient failure in servo system testing

Computer based data conditioning and display technique detects incipient failure in servo system testing, for use in prelaunch checkout of complex nonlinear servomechanisms. These phase plane displays enable identification of, on line, unusual or abnormal servo responses which can be displayed compactly in the time domain on a cathode ray tube.

Affenito, F. J.↗

Super-Zip separation joint performance investigation

Following functional failures of two of five Lockheed Super-Zip spacecraft separation joints in a development test series to quantify thermal effects, an investigation program was initiated on this and related systems to assist in preventing recurrence. The Super-Zip joint, applied in a ring configuration, utilizes an explosively expanded tube to fracture surrounding prenotched aluminum plates to achieve planar separation. A unique test method was developed and more than 300 individual test firings were conducted to provide an understanding of severance mechanisms, the functional performance effects of system variables, and the most likely cause of system failure. An approach for defining functional margin was developed, as well as specific recommendations for improving existing and future systems.

Bement, Laurence J.↗

Person to Person Biological Heat Bypass During EVA Emergencies

During EVA and other extreme environments, mutual human support is sometimes the last way to survive when there is a failure of the life support equipment. The possibility to transfer a coolant to remove heat or a warming fluid to increase heat from one individual to another to support the thermal balance of the individual with system failure was assessed. The following scenarios were considered: 1. one participant has a cooling system that is not working well and already has a body heat deficit equal to 100-120 kcal and a finger temperature decline to 25 C; 2. one participant has the same status of overcooling and the other mild overheating. Preliminary findings showed promise in using such sharing tactics to extend the time duration of survival in extreme situations when there is a high metabolic rate in the donor.

Koscheyev, Victor S.↗

Kivalina Biomass Reactor

This report summarizes work performed under DOE Award DE-EE00010149 to support the reliable operation of a community-scale biochar reactor system in Kivalina, Alaska. The project focused on improving sanitation and waste management in a remote community by assessing the installed system, identifying spare parts, defining key performance indicators (KPIs), preparing operator and maintenance manuals, and developing mobile reporting tools for operational data and KPI tracking. The team also produced training materials and recorded videos to support operator onboarding and continuity. The project demonstrated progress in system readiness, documentation, and digital reporting, while also identifying challenges common to remote deployments, including travel constraints, upstream system failures, and local resource limitations. This work provides a practical framework for improving the operation, monitoring, and future replication of biomass reactor systems in remote communities.

09 BIOMASS FUELS↗

Application of system safety to rail transit systems

Management emphasis on system safety in the rapid transit industry includes the granting and use of funds by the Federal Government according to systematic analysis of safety hazards in advance. Likelihood predictions that those hazards will be activated by exposure of the system to a system failure, a human error, external conditions, or combinations of these aspects determine alternatives to the assumption of risk and recommend corrections before the system is operational. Rigorous safety analyses are projected to assure operational safety for prolonged periods under varied maintenance conditions; these analysis encompass station accident possibilities as well as train-person collisions, car equipment and design, traffic control systems, and tunnel design problems.

Thomas DeW. Styles↗