Search NASA⌕ Search

SEARCH · Search NASA

Results for “fail operational”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 181 records · Page 10

Heroic Reliability Improvement in Manned Space Systems

System reliability can be significantly improved by a strong continued effort to identify and remove all the causes of actual failures. Newly designed systems often have unexpected high failure rates which can be reduced by successive design improvements until the final operational system has an acceptable failure rate. There are many causes of failures and many ways to remove them. New systems may have poor specifications, design errors, or mistaken operations concepts. Correcting unexpected problems as they occur can produce large early gains in reliability. Improved technology in materials, components, and design approaches can increase reliability. The reliability growth is achieved by repeatedly operating the system until it fails, identifying the failure cause, and fixing the problem. The failure rate reduction that can be obtained depends on the number and the failure rates of the correctable failures. Under the strong assumption that the failure causes can be removed, the decline in overall failure rate can be predicted. If a failure occurs at the rate of lambda per unit time, the expected time before the failure occurs and can be corrected is 1/lambda, the Mean Time Before Failure (MTBF). Finding and fixing a less frequent failure with the rate of lambda/2 per unit time requires twice as long, time of 1/(2 lambda). Cutting the failure rate in half requires doubling the test and redesign time and finding and eliminating the failure causes.Reducing the failure rate significantly requires a heroic reliability improvement effort.

life support↗

The SMAP Mission Combined Active-Passive Soil Moisture Product at 9 Km and 3 Km Spatial Resolutions

The NASA Soil Moisture Active Passive (SMAP) mission was launched on January 31st, 2015. The spacecraft was to provide high-resolution (3 km and 9 km) global soil moisture estimates at regular intervals by combining for the first time L-band radiometer and radar observations. On July 7th, 2015, a component of the SMAP radar failed and the radar ceased operation. However, before this occurred the mission was able to collect and process ~2.5 months of the SMAP high-resolution active-passive soil moisture data (L2SMAP) that coincided with the Northern Hemisphere's vegetation green-up and crop growth season. In this study, we evaluate the SMAP high-resolution soil moisture product derived from several alternative algorithms against in situ data from core calibration and validation sites (CVS), and sparse networks. The baseline algorithm had the best comparison statistics against the CVS and sparse networks. The overall unbiased root-mean-square-difference is close to the 0.04 cu. m/cu. m the SMAP mission requirement. A 3 km spatial resolution soil moisture product was also examined. This product had an unbiased root-mean-square-difference of ~0.053 cu. m/cu. m. The SMAP L2SMAP product for ~2.5 months is now validated for use in geophysical applications and research and available to the public through the NASA Distributed Active Archive Center (DAAC) at the National Snow and Ice Data Center (NSIDC). The L2SMAP product is packaged with the geo-coordinates, acquisition times, and all requisite ancillary information. Although limited in duration, SMAP has clearly demonstrated the potential of using a combined L-band radar-radiometer for proving high spatial resolution and accurate global soil moisture.

high resolution soil moisture↗

Stirling Engine Controller

Stirling technology is being developed to replace RTG s (Radioisotope Thermoelectric Generators), more specifically a stirling convertor, which is a stirling engine coupled to a linear alternator. Over the past three decades, the stirling engine has been designed to perform different functions. Stirling convertors have been designed to decrease fuel consumption in automobiles. They have also been designed for terrestrial and space applications. Currently NASA Glenn is using the convertor for space based applications. A stiring converter is a better means of power for deep space mission and "dusty" mission, like the Mars Rovers, than solar panels because it is not affected by dust. Spirit and Opportunity, two Mars rovers currently navigating the planet, are losing their ability to generate electricity because dust is collecting on their solar panels. Opportunity is losing more energy because its robotic arm has a heater with a switch that can not be turned off. The heater is not needed at night, but yet still runs. This generates a greater loss of electricity and in turn diminishes the performance of the rover. The stirling cycle has the potential to provide very efficient conversion of heat energy to electric a1 energy, more so than RTG's. The stirling engine converts the thermal energy produced by the decaying radioisotope to mechanical energy; the linear alternator converts this into electricity. convertor. Since the early 1990's tests have been performed to maximize the efficiency of the stirling converter. Many months, even years, are dedicated to preparing and performing tests. Currently, two stirling convertors #'s 13 and 14, which were developed by Stirling Technology Company, are on an extended operation test. As of June 7th, the two convertors reached 7,500 hours each of operation. Before the convertors could run unattended, many safety precautions had to be examined. So, special instrumentation and circuits were developed to detect off nominal conditions and also safely shutdown the engines. The test will last for a period of 8000 to 9000 hours. Other types of tests that have been performed are: performance mapping, controller development, launch environment, and vibration emissions testing. Currently, the thermo-mechanical system branch is housing a RG-350, a stirling convertor. The convertor was used in previous tests such as a Hall Thruster test, world s first integrated test of a dynamic power system with electric propulsion. Another test performed was to conclude if free piston stirling convertors can be synchronized for vibration balancing, with no thermodynamic or electrical connections and not cause both to shutdown if one failed. The ability to reduce vibration by synchronizing convertor operation but still be able to operate when one partner fails is pertinent in space and terrestrial applications. The convertor is now being brought back into operation and a controller is in the process of being developed. This convertor will be used as a testbed for new controllers. I worked with Mary Ellen Roth on the electric engineering aspects of the RG-350. My main goal was to enhance the data collection process. I worked on different aspects of the RG-350, with a main focus on the engine controller. I drew a schematic of the wire connections in the engine controller, using PCB Express, so that a plan could be devised to connect the power meter properly between the output of the engine and the engine controller. I measured the power using two different instruments: Valhalla Scientific power meter and Ohio Semitronics power measurement device. The convertor is connected to an Agilent 34970A Data Acquisition/Switch Unit, which allows the user to measure, record, and monitor voltage, current, frequency, and temperature. I assisted in preparing the Data Acquisition for general operation. I also helped test a panel of transducers, which will be placed in the rack that powers and monitors the convertor.

Blaze, Gina M.↗

Cryogenic Two-Phase Flight Experiment: Results overview

This paper focuses on the flight results of the Cryogenic Two-Phase Flight Experiment (CRYOTP), which was a Hitchhiker based experiment that flew on the space shuttle Columbia in March of 1994 (STS-62). CRYOTP tested two new technologies for advanced cryogenic thermal control; the Space Heat Pipe (SHP), which was a constant conductance cryogenic heat pipe, and the Brilliant Eyes Thermal Storage Unit (BETSU), which was a cryogenic phase-change thermal storage device. These two devices were tested independently during the mission. Analysis of the flight data indicated that the SHP was unable to start in either of two attempts, for reasons related to the fluid charge, parasitic heat leaks, and cryocooler capacity. The BETSU test article was successfully operated with more than 250 hours of on-orbit testing including several cooldown cycles and 56 freeze/thaw cycles. Some degradation was observed with the five tactical cryocoolers used as thermal sinks, and one of the cryocoolers failed completely after 331 hours of operation. Post-flight analysis indicated that this problem was most likely due to failure of an electrical controller internal to the unit.

Swanson, T.↗

Thermionic cathode life-test studies

A NASA-Lewis Research Center program for life testing commercial, high-current-density thermionic cathodes has been in progress since 1971. The purpose of the program is to develop long-life power microwave tubes for space communications. Four commercial-type cathodes are being evaluated in this investigation. They are the 'Tungstate', 'S' type, 'B' type, and 'M' type cathodes, all of which are capable of delivering 1 A/ sq cm or more of emission current at an operating temperature in the range of 1000-1100 C. The life test vehicles used in these studies are similar in construction to that of a high-power microwave tube and employ a high-convergence electron-gun structure; in contrast to earlier studies that used close-space diodes. These guns were designed for operation at 2 A/sq cm of cathode loading. The 'Tungstate' cathodes failed at 700 h or less and the 'S' cathode exhibited a lifetime of about 20,000 h. One 'B' cathode has failed after 27,000 h, the remaining units continuing to operate after up to 30,000 h. Only limited data are now available for the 'M' cathode, because only one has been operated for as long as 19,000 h. However, the preliminary results indicate the emission current from the 'M' cathode is more stable than the 'B' cathode and that it can be operated at a true temperature approximately 100 C lower than for the 'B' cathode.

Forman, R.↗

SAMPEX Spin Stabilized Mode

The Solar, Anomalous, and Magnetospheric Particle Explorer (SAMPEX), the first of the Small Explorer series of spacecraft, was launched on July 3, 1992 into an 82' inclination orbit with an apogee of 670 km and a perigee of 520 km and a mission lifetime goal of 3 years. After more than 15 years of continuous operation, the reaction wheel began to fail on August 18,2007. With a set of three magnetic torquer bars being the only remaining attitude actuator, the SAMPEX recovery team decided to deviate from its original attitude control system design and put the spacecraft into a spin stabilized mode. The necessary operations had not been used for many years, which posed a challenge. However, on September 25, 2007, the spacecraft was successfully spun up to 1.0 rpm about its pitch axis, which points at the sun. This paper describes the diagnosis of the anomaly, the analysis of flight data, the simulation of the spacecraft dynamics, and the procedures used to recover the spacecraft to spin stabilized mode.

Tsai, Dean C.↗

Enhanced Component Performance Study: Motor Driven Pumps 1998-2024

This report presents an enhanced performance evaluation of motor-driven pumps (MDPs) at U.S. commercial nuclear power plants. The data used in this study are based on the operating experience failure reports from calendar year 1998 through 2024 as reported in the Institute of Nuclear Power Operations (INPO) Industry Reporting and Information System (IRIS). The MDP failure modes considered for standby systems are fail to start (FTS), fail to run (FTR) for one hour of operation (FTR=1H), FTR after one hour of operation (FTR>1H), and for normally running systems FTS and FTR. An eight-hour unreliability estimate is also calculated and trended. The component reliability estimates and the reliability data are trended for the most recent 10-year period while yearly estimates for reliability are provided for the entire study period.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Enhanced Component Performance Study: Turbine-Driven Pumps 1998-2024

This report presents an enhanced performance evaluation of turbine driven pumps (TDPs) at U.S. commercial nuclear power plants. The data used in this study are based on the operating experience failure reports from calendar year 1998 through 2024 as reported in the Institute of Nuclear Power Operations (INPO) Industry Reporting and Information System (IRIS). The TDP failure modes considered for standby systems are fail to start (FTS), fail to run (FTR) for one hour of operation (FTR=1H), FTR after one hour of operation (FTR>1H), and for normally running systems FTS and FTR. An eight hour unreliability estimate is also calculated and trended. The component reliability estimates and the reliability data are trended for the most recent 10 year period while yearly estimates for reliability are provided for the entire study period. No increasing trends were identified for TDPs for the most recent 10 year period.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Effects of Autonomous sUAS Separation Methods on Subjective Workload, Situation Awareness, and Trust

The Unmanned Aircraft System (UAS) Traffic Management (UTM) concept was designed to support autonomous small UAS operations at a large-scale and without direct human intervention. However, human-autonomy interactions will be impacted by situation awareness, workload, and trust in the autonomy. Method: Nine participants monitored live small UAS operations in a representative UTM system during a series of traffic conflict scenarios and then provided subjective responses regarding situation awareness, workload, and trust in the autonomous separation method. The study employed a 3 (Separation Method: Autonomous Sense and Avoid, Geofence, Manual) × 2 (Incursion: High, Medium) within subjects design. Results: Situation awareness ratings for both autonomous separation methods were significantly lower than the manual condition. An interaction indicated differential workload ratings for the Autonomous Sense and Avoid separation ratings. Trust ratings significantly dropped when the Geofencing separation method failed. Conclusion: Subjective responses of remote operators in the UTM system are affected by the vehicle separation methods. Operators’ understanding of decisions made by the autonomous systems onboard the vehicle likely influence this effect

UAS↗

NextGen Flight Deck Surface Trajectory-Based Operations (STBO): Contingency Holds

The purpose of this pilot-in-the-loop taxi simulation was to investigate a NextGen Surface Trajectory-Based Operations (STBO) concept called "contingency holds." The contingency-hold concept parses a taxi route into segments, allowing an air traffic control (ATC) surface traffic management (STM) system to hold an aircraft when necessary for safety. Under nominal conditions, if the intersection or active runway crossing is clear, the hold is removed, allowing the aircraft to continue taxiing without slowing, thus improving taxi efficiency, while minimizing the excessive brake use, fuel burn, and emissions associated with stop-and-go taxi. However, when a potential traffic conflict exists, the hold remains in place as a fail-safe mechanism. In this departure operations simulation, the taxi clearance included a required time of arrival (RTA) to a specified intersection. The flight deck was equipped with speed-guidance avionics to aid the pilot in safely meeting the RTA. On two trials, the contingency hold was not released, and pilots were required to stop. On two trials the contingency hold was released 15 sec prior to the RTA, and on two trials the contingency hold was released 30 sec prior to the RTA. When the hold remained in place, all pilots complied with the hold. Results also showed that when the hold was released at 15-sec or 30-sec prior to the RTA, the 30-sec release allowed pilots to maintain nominal taxi speed, thus supporting continuous traffic flow; whereas, the 15-sec release did not. The contingency-hold concept, with at least a 30-sec release, allows pilots to improve taxiing efficiency by reducing braking, slowing, and stopping, but still maintains safety in that no pilots "busted" the clearance holds. Overall, the evidence suggests that the contingency-hold concept is a viable concept for optimizing efficiency while maintaining safety.

NextGen↗

SCOAPE-II: A 2024 Multiplatform Measurement Campaign off the US Gulf Coast to Assess Oil and Gas Emissions on the Outer Continental Shelf

Nine years ago, the Department of Interior’s Bureau of Ocean Energy Management (BOEM), the Agency with Air Quality (AQ) jurisdiction over the Outer Continental Shelf (OCS) of the US Gulf Coast west of 87.5° W longitude, asked NASA to determine the feasibility of using satellite data to measure offshore emissions in a region of concentrated oil and natural gas (ONG) operations. To study this issue NASA and BOEM conducted the May 2019 Satellite Coastal and Oceanic Atmospheric Pollution Experiment (SCOAPE) cruise in the Gulf. SCOAPE addressed both technological and scientific issues related to measuring nitrogen dioxide (NO 2 , a common air pollutant), including contrasting near-shore and deepwater regimes. Given the April 2023 launch of the geostationary Tropospheric Emissions: Monitoring of Pollution (TEMPO) AQ satellite, a 2024 SCOAPE-II was conducted in the Gulf with both ship and aircraft measurements. We present an overview of the SCOAPE-II campaign, analysis and validation of satellite-observed NO 2 , and evaluate measurements of methane from ship, aircraft, and satellite near ONG platforms. Our SCOAPE-II results are as follows: 1) Satellite NO 2 measurements (∼13:30 local time) from the TROPOspheric Monitoring Instrument (TROPOMI) are more accurate than TEMPO’s hourly scans (8.6% vs. 23.6% mean absolute bias); a new version of TEMPO data is currently being processed; 2) ship and aircraft measurements captured dozens of NO 2 and methane plumes from ONG operations, showing that they are persistent emitters; 3) satellite measurements of methane failed to replicate ship and aircraft measurements, presenting ongoing challenges for operational emissions monitoring over the Gulf.

satellite validation↗

Space Transportation System Availability Relationships to Life Cycle Cost

Future space transportation architectures and designs must be affordable. Consequently, their Life Cycle Cost (LCC) must be controlled. For the LCC to be controlled, it is necessary to identify all the requirements and elements of the architecture at the beginning of the concept phase. Controlling LCC requires the establishment of the major operational cost drivers. Two of these major cost drivers are reliability and maintainability, in other words, the system's availability (responsiveness). Potential reasons that may drive the inherent availability requirement are the need to control the number of unique parts and the spare parts required to support the transportation system's operation. For more typical space transportation systems used to place satellites in space, the productivity of the system will drive the launch cost. This system productivity is the resultant output of the system availability. Availability is equal to the mean uptime divided by the sum of the mean uptime plus the mean downtime. Since many operational factors cannot be projected early in the definition phase, the focus will be on inherent availability which is equal to the mean time between a failure (MTBF) divided by the MTBF plus the mean time to repair (MTTR) the system. The MTBF is a function of reliability or the expected frequency of failures. When the system experiences failures the result is added operational flow time, parts consumption, and increased labor with an impact to responsiveness resulting in increased LCC. The other function of availability is the MTTR, or maintainability. In other words, how accessible is the failed hardware that requires replacement and what operational functions are required before and after change-out to make the system operable. This paper will describe how the MTTR can be equated to additional labor, additional operational flow time, and additional structural access capability, all of which drive up the LCC. A methodology will be presented that provides the decision makers with the understanding necessary to place constraints on the design definition. This methodology for the major drivers will determine the inherent availability, safety, reliability, maintainability, and the life cycle cost of the fielded system. This methodology will focus on the achievement of an affordable, responsive space transportation system. It is the intent of this paper to not only provide the visibility of the relationships of these major attribute drivers (variables) to each other and the resultant system inherent availability, but also to provide the capability to bound the variables, thus providing the insight required to control the system's engineering solution. An example of this visibility is the need to provide integration of similar discipline functions to allow control of the total parts count of the space transportation system. Also, selecting a reliability requirement will place a constraint on parts count to achieve a given inherent availability requirement, or require accepting a larger parts count with the resulting higher individual part reliability requirements. This paper will provide an understanding of the relationship of mean repair time (mean downtime) to maintainability (accessibility for repair), and both mean time between failure (reliability of hardware) and the system inherent availability.

Rhodes, Russel E.↗

Sixty-four-Channel Inline Cable Tester

Faults in wiring are a serious concern for the aerospace and aeronautics (commercial, military, and civil) industries. A number of accidents have occurred because faulty wiring created shorts or opens that resulted in the loss of control of the aircraft or because arcing led to fires and explosions. Some of these accidents have resulted in the massive loss of lives (such as in the TWA Flight 800 accident). Circuits on the Space Shuttle have also failed because of faulty insulation on wiring. STS-93 lost power when a primary power circuit in one engine failed and a second engine had a backup power circuit fault. Cables are usually tested on the ground after the crew reports a fault encountered during flight. Often such failures result from vibration and cannot be replicated while the aircraft is stationary. It is therefore important to monitor faults while the aircraft is in operation, when cables are more likely to fail. Work is in progress to develop a cable fault tester capable of monitoring up to 64 individual wires simultaneously. Faults can be monitored either inline or offline. In the inline mode of operation, the monitoring is performed without disturbing the normal operation of the wires under test. That is, the operations are performed unintrusively and are essentially undetectable for the test signal levels are below the noise floor. A cable can be monitored several times per second in the offline mode and once a second in the inline mode. The 64-channel inline cable tester not only detects the occurrence of a fault, but also determines the type of fault (short/open) and the location of the fault. This will enable the detection of intermittent faults that can be repaired before they become serious problems.

Source record↗

Dynamic Temporal Graph Sequence Data for Resilience-Oriented Distribution Network Reconfiguration

This dataset comprises temporal dynamic graph sequences generated from power grid simulations focused on grid reconfiguration to enhance resilience. The simulations model failure propagation under varying conditions, with nodes assigned distinct failure probabilities. For each time step, the dataset captures the evolution of node states (functional or failed) and features critical to grid operations, such as pv_output, load_profile, load_dispatch, dg_output, loss, and voltage. Node types include sources, normal loads, and nodes with specific equipment like PVs, micro turbines, or shunt capacitors. The dataset is structured to support the training of dynamic graph neural networks, facilitating research on node feature prediction and edge dynamics under failure scenarios. Three distinct configurations are included, providing a robust foundation for modeling power grid resilience.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Product Defect Detection System: SYSM- 5620 Final Project

Retail sales is a growing market estimated to up to seven percent year over year. With this growing market there is also a trend in growing rate of retail returns, estimated just last year at $\$$850 billion. Retail stores must ensure that products available for purchase remain safe, undamaged, and acceptable to customers throughout their time in the store. This job exists regardless of the specific solution used because stores are always responsible for preventing damaged or defective products from reaching customers and when they fail to this is categorized under operation inefficiencies which accounts for an estimated $\$$12 billion in returns. When defective items remain on the sales floor, stores may experience increased returns, reduced customer satisfaction, loss of customer trust, and potential safety concerns depending on the product type. As a result, the core job to be done is to identify defective products quickly, remove them from the sales floor before they are purchased, and preserve useful information about the defect so that the store can improve its handling, stocking, and supplier coordination over time. The need for a more reliable process is especially important in high volume retail environments where employees manage large numbers of products across many aisles, shelves, and storage areas. In these settings, manual inspection alone can be inconsistent and difficult to sustain at the individual item level. At the same time, broader retail trends continue to emphasize operational efficiency, product visibility, and improved customer experience, creating an opportunity for more automated and data driven defect detection methods.

42 ENGINEERING↗

Integrated orbital servicing and payloads study. Volume 1: Executive summary

A study is summarized in which a comparison was made of the following modes of maintaining a satellite system: (1) expendable mode in which failed satellites are replaced, (2) on-orbit servicing where a satellite can be fixed by unmanned module exchange in space, and (3) ground refurbishment in which the satellite is brought back to ground for repairs. It was concluded that on-orbit maintenance is the most cost-effective mode and that it is technically feasible. It can be used to repair failed satellites, to improve reliability of operating satellites, and to update equipment. On-orbit servicing can increase program flexibility and satellite reliability, lifetime, and availability. The significant conclusions and results of two studies are summarized.

Source record↗

Simulator study of the low-speed handling qualities of a supersonic cruise arrow-wing transport configuration during approach and landing

A fixed-based simulator study was conducted to determine the low-speed flight characteristics of an advanced supersonic cruise transport having an arrow wing, a horizontal tail, and four dry turbojets with variable geometry turbines. The primary piloting task was the approach and landing. The statically unstable (longitudinally) subject configuration has unacceptable low-speed handling qualities with no augmentation. Therefore, a hardened stability augmentation system is required to achieve acceptable handling qualities, should the normal operational stability and control augmentation system fail. In order to achieve satisfactory handling qualities, considerable augmentation was required.

Grantham, W. D.↗

Reliability analysis of forty-five strain-gage systems mounted on the first fan stage of a YF-100 engine

The reliability of 45 state-of-the-art strain gage systems under full scale engine testing was investigated. The flame spray process was used to install 23 systems on the first fan rotor of a YF-100 engine; the others were epoxy cemented. A total of 56 percent of the systems failed in 11 hours of engine operation. Flame spray system failures were primarily due to high gage resistance, probably caused by high stress levels. Epoxy system failures were principally erosion failures, but only on the concave side of the blade. Lead-wire failures between the blade-to-disk jump and the control room could not be analyzed.

Holanda, R.↗