Search NASA⌕ Search

SEARCH · Search NASA

Results for “Software Failures”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 559 records · Page 31

Fault Injection Campaign for a Fault Tolerant Duplex Framework

Fault tolerance is an efficient approach adopted to avoid or reduce the damage of a system failure. In this work we present the results of a fault injection campaign we conducted on the Duplex Framework (DF). The DF is a software developed by the UCLA group [1, 2] that uses a fault tolerant approach and allows to run two replicas of the same process on two different nodes of a commercial off-the-shelf (COTS) computer cluster. A third process running on a different node, constantly monitors the results computed by the two replicas, and eventually restarts the two replica processes if an inconsistency in their computation is detected. This approach is very cost efficient and can be adopted to control processes on spacecrafts where the fault rate produced by cosmic rays is not very high.

fault injector↗

Connecting Research and Practice: An Experience Report on Research Infusion with SAVE

NASA systems need to be highly dependable to avoid catastrophic mission failures. This calls for rigorous engineering processes including meticulous validation and verification. However, NASA systems are often highly distributed and overwhelmingly complex, making the software portion of these systems challenging to understand, maintain, change, reuse, and test. NASA's systems are long-lived and the software maintenance process typically constitutes 60-80% of the total cost of the entire lifecycle. Thus, in addition to the technical challenges of ensuring high life-time quality of NASA's systems, the post-development phase also presents a significant financial burden. Some of NASA's software-related challenges could potentially be addressed by some of the many powerful technologies that are being developed in software research laboratories. Many of these research technologies seek to facilitate maintenance and evolution by for example architecting, designing and modeling for quality, flexibility, and reuse. Other technologies attempt to detect and remove defects and other quality issues by various forms of automated defect detection, architecture analysis, and various forms of sophisticated simulation and testing. However promising, most such research technologies nevertheless do not make the transition from the research lab to the software lab. One reason the transition from research to practice seldom occurs is that research infusion and technology transfer is difficult. For example, factors related to the technology are sometimes overshadowed by other types of factors such as reluctance to change and therefore prohibits the technology from sticking. Successful infusion might also take very long time. One famous study showed that the discrepancy between the conception of the idea and its practical use was 18 years plus or minus three. Nevertheless, infusing new technology is possible. We have found that it takes special circumstances for such research infusion to succeed: 1) there must be evidence that the technology works in the practitioner's particular domain, 2) there must be a potential for great improvements and enhanced competitive edge for the practitioner, 3) the practitioner has to have strong individual curiosity and continuous interest in trying out new technologies, 4) the practitioner has to have support on multiple levels (i.e. from the researchers, from management, from sponsors etc), and 5) to remain infused, the new technology has to be integrated into the practitioner's processes so that it becomes a natural part of the daily work. NASA IV&V's Research Infusion initiative sponsored by NASA's Office of Safety & Mission Assurance (OSMA) through the Software Assurance Research Program (SARP), strives to overcome some of the problems related to research infusion.

Lindvall, Mikael↗

"Making Safety Happen" Through Probabilistic Risk Assessment at NASA

NASA is using Probabilistic Risk Assessment (PRA) as one of the tools in its Safety & Mission Assurance (S&MA) tool belt to identify and quantify risks associated with human spaceflight. This paper discusses some of the challenges and benefits associated with developing and using PRA for NASA human space programs. Some programs have entered operation prior to developing a PRA, while some have implemented PRA from the start of the program. It has been observed that the earlier a design change is made in the concept or design phase, the less impact it has on cost and schedule. Not finding risks until the operation phase yields much costlier design changes and major delays, which can result in discussions of just accepting the risk. Risk contributors identified by PRA are not just associated with hardware failures. They include but are not limited to crew fatality due to medical causes, the environment the vehicle and crew are exposed to, the software being used, and the reliability of the crew performing required actions. Some programs have entered operation prior to developing a PRA, and while PRA can still provide a benefit for operations and future design trades, the benefit of implementing PRA from the start of the program provides the added benefit of informing design and reducing risk early in program development. Currently, NASA’s International Space Station (ISS) program is in its 20th year of on-orbit operations around the Earth and has several new programs in the design phase preparing to enter the operation phase all of which have active (or living) PRAs. These programs incorporate PRA as part of their Risk-Informed, Decision-Making (RIDM) process. For new NASA human spaceflight programs discussion begins with mission concept, establishing requirements, forming the PRA team, and continues through the design cycles into the operational phase. Several examples of PRA related applications and observed lessons are included.

Applications↗

Designing and Implementing a Distributed System Architecture for the Mars Rover Mission Planning Software (Maestro)

Distributed systems allow scientists from around the world to plan missions concurrently, while being updated on the revisions of their colleagues in real time. However, permitting multiple clients to simultaneously modify a single data repository can quickly lead to data corruption or inconsistent states between users. Since our message broker, the Java Message Service, does not ensure that messages will be received in the order they were published, we must implement our own numbering scheme to guarantee that changes to mission plans are performed in the correct sequence. Furthermore, distributed architectures must ensure that as new users connect to the system, they synchronize with the database without missing any messages or falling into an inconsistent state. Robust systems must also guarantee that all clients will remain synchronized with the database even in the case of multiple client failure, which can occur at any time due to lost network connections or a user's own system instability. The final design for the distributed system behind the Mars rover mission planning software fulfills all of these requirements and upon completion will be deployed to MER at the end of 2005 as well as Phoenix (2007) and MSL (2009).

Goldgof, Gregory M.↗

Lessons Learned in the Livingstone 2 on Earth Observing One Flight Experiment

The Livingstone 2 (L2) model-based diagnosis software is a reusable diagnostic tool for monitoring complex systems. In 2004, L2 was integrated with the JPL Autonomous Sciencecraft Experiment (ASE) and deployed on-board Goddard's Earth Observing One (EO-1) remote sensing satellite, to monitor and diagnose the EO-1 space science instruments and imaging sequence. This paper reports on lessons learned from this flight experiment. The goals for this experiment, including validation of minimum success criteria and of a series of diagnostic scenarios, have all been successfully net. Long-term operations in space are on-going, as a test of the maturity of the system, with L2 performance remaining flawless. L2 has demonstrated the ability to track the state of the system during nominal operations, detect simulated abnormalities in operations and isolate failures to their root cause fault. Specific advances demonstrated include diagnosis of ambiguity groups rather than a single fault candidate; hypothesis revision given new sensor evidence about the state of the system; and the capability to check for faults in a dynamic system without having to wait until the system is quiescent. The major benefits of this advanced health management technology are to increase mission duration and reliability through intelligent fault protection, and robust autonomous operations with reduced dependency on supervisory operations from Earth. The work-load for operators will be reduced by telemetry of processed state-of-health information rather than raw data. The long-term vision is that of making diagnosis available to the onboard planner or executive, allowing autonomy software to re-plan in order to work around known component failures. For a system that is expected to evolve substantially over its lifetime, as for the International Space Station, the model-based approach has definite advantages over rule-based expert systems and limit-checking fault protection systems, as these do not scale well. The model-based approach facilitates reuse of the L2 diagnostic software; only the model of the system to be diagnosed and telemetry monitoring software has to be rebuilt for a new system or expanded for a growing system. The hierarchical L2 model supports modularity and expendability, and as such is suitable solution for integrated system health management as envisioned for systems-of-systems.

Hayden, Sandra C.↗

International Space Station Common Cabin Air Assembly Water Separator On-Orbit Operation, Failure, and Redesign

The ability to control the temperature and humidity of an environment or habitat is critical for human survival. These factors are important to maintaining human health and comfort, as well as maintaining mechanical and electrical equipment in good working order to support the human and to accomplish mission objectives. The temperature and humidity of the International Space Station (ISS) United States On-orbit Segment (USOS) cabin air is controlled by the Common Cabin Air Assembly (CCAA). The CCAA consists of a fan, a condensing heat exchanger (CHX), an air/water separator, temperature and liquid sensors, and electrical controlling hardware and software. The Water Separator (WS) pulls in air and water from the CHX, and centrifugally separates the mixture, sending the water to the condensate bus and the air back into the CHX outlet airstream. Two distinct early failures of the CCAA Water Separator in the Quest Airlock forced operational changes and brought about the re-design of the Water Separator to improve the useful life via modification kits. The on-orbit operational environment of the Airlock presented challenges that were not foreseen with the original design of the Water Separator. Operational changes were instituted to prolong the life of the third installed WS, while waiting for newly designed Water Separators to be delivered on-orbit. The modification kit design involved several different components of the Water Separator, including the innovative use of a fabrication technique to build the impellers used in Water Separators out of titanium instead of aluminum. The technique allowed for the cost effective production of the low quantity build. This paper will describe the failures of the Water Separators in the Quest Airlock, the operational constraints that were implemented to prolong the life of the installed Water Separators throughout the USOS, and the innovative re-design of the CCAA Water Separator.

Balistreri, Steven F., Jr.↗

Success Path Method: Introduction to the Success Path Method Software Tool©

As part of its commitment to advancing safety and reliability assessment methodologies, Argonne National Laboratory pioneered the use of an evaluation method called the Success Path Method (SPM) to improve risk management for offshore oil and gas operations. The development of the SPM at Argonne has been driven by the need to improve existing risk assessment methodologies by focusing on the steps necessary for success rather than failure modes alone. This is particularly important for industrial environments like offshore facilities that perform multiple functions under a continuously evolving set of operational conditions – such as water depth and temperature, currents, and weather conditions. In these dynamic environments, the traditional Probabilistic Risk Assessment (PRA) approach is far too complex as it focuses on what can go wrong – which comprises an infinite failure space that must be fully explored and understood. By shifting the focus to a finite space of success paths, the SPM enables operators and decision makers to prioritize a manageable number of steps that must go right to ensure success. Building on its five decades of experience in safety assessments for the nuclear industry, Argonne made major adaptations to existing risk assessment methods utilizing features similar to fault trees that are traditionally used in PRA to map all pathways in which the system can malfunction. In contrast, SPM identifies the components and processes that must function correctly to achieve specific outcomes – such as preventing the uncontrolled release of hydrocarbons during drilling operations. The SPM framework integrates equipment, procedures, software, processes, and human actions to ensure that physical barriers meet critical safety functions in dynamic operational conditions. This approach helps identify failure modes and improve operational risk management by narrowing the focus to key success elements, which in turn reduces uncertainty and helps users understand, manage, and respond to failures.

97 MATHEMATICS AND COMPUTING↗

Computer aided control of a mechanical arm

A method for computer-aided remote control of a six-degree-of-freedom manipulator arm involved in the on-orbit servicing of a spacecraft is presented. The control configuration features a supervisory type of control in which each of the segments of a module exchange trajectory is controlled automatically under human supervision, with manual commands to proceed to the next step and in the event of a failure or undesirable outcome. The implementation of the supervisory system is discussed in terms of necessary onboard and ground- or Orbiter-based hardware and software, and a one-g demonstration system built to allow further investigation of system operation is described. Possible applications of the system include the construction of satellite solar power systems, environmental testing and the control of heliostat solar power stations.

Derocher, W. L., Jr.↗

Optimal platform skewing for Space Shuttle inertial measurement unit redundancy management

Constraints are applied to a general quaternion which describes the skewing between platforms of the Space Shuttle IMU. Once a skewing is derived, the use of the failure magnitude to threshold ratio makes possible predictions of the identification sensitivities for various failure modes. This in turn simplifies analyses and identifies portions of the flight envelope where second failure coverage is lacking. The square root of 6 and square root of 2 skewings have been baselined for use during nominal entry; the realignment software will be used on orbit to reskew the IMUs to the optimal configuration.

Rasmussen, M. C.↗

Engineering challenges of in-flight spacecraft - Voyager: A case study

Some of the engineering problems encountered during the post-launch phase of interplanetary space missions are described, with emphasis given to the Voyager missions. The major in-flight modifications in Voyager spacecraft's operational capability with respect to communications, payload, and navigation systems are discussed. Attention is given to the instances of 'failure workaround' including: recovery from a failed receiver, recovery from a seized scan platform actuator, and modifications to the Attitude Articulation and Control Subsystem (AACS) software during the Saturn encounter. A detailed line drawing of the Voyager spacecraft is provided.

Jones, C. P.↗

How to Extend the Capabilities of Space Systems for Long Duration Space Exploration Systems

For sustainable Exploration Missions the need exists to assemble systems-of-systems in space, on the Moon or on other planetary surfaces. To fulfill this need new and innovative system architectures must be developed to be modularized and launched with the present lift capability of existing rocket technology. To enable long duration missions with minimal redundancy and mass, system software and hardware must be reconfigurable. This will enable increased functionality and multiple use of launched assets while providing the capability to quickly overcome components failures. Additional required capability includes the ability to dynamically demate and reassemble individual system elements during a mission in order to recover from failed hardware or to adapt to changes in mission requirements. To meet the Space Exploration goals of Interoperability and Reconfigurability, many challenges must be addressed to transform the traditional static avionics architectures into architectures with dynamic capabilities. The objective of this paper is to introduce concepts associated with reconfigurable computer systems; to review the various needs and challenges associated with reconfigurable avionics space systems; to provide an operational example that illustrates the application to both the Crew Exploration Vehicle and a collection of 'Habot-like' mobile surface elements; to summarize the approaches that address key challenges to the acceptance of a Flexible, Intelligent, Modular, Affordable and Reconfigurable avionics space system.

space assembly↗

Models Extracted from Text for System-Software Safety Analyses

This presentation describes extraction and integration of requirements information and safety information in visualizations to support early review of completeness, correctness, and consistency of lengthy and diverse system safety analyses. Software tools have been developed and extended to perform the following tasks: 1) extract model parts and safety information from text in interface requirements documents, failure modes and effects analyses and hazard reports; 2) map and integrate the information to develop system architecture models and visualizations for safety analysts; and 3) provide model output to support virtual system integration testing. This presentation illustrates the methods and products with a rocket motor initiation case.

Malin, Jane T.↗

Multiscale Fatigue Life Prediction for Composite Panels

Fatigue life prediction capabilities have been incorporated into the HyperSizer Composite Analysis and Structural Sizing Software. The fatigue damage model is introduced at the fiber/matrix constituent scale through HyperSizer s coupling with NASA s MAC/GMC micromechanics software. This enables prediction of the micro scale damage progression throughout stiffened and sandwich panels as a function of cycles leading ultimately to simulated panel failure. The fatigue model implementation uses a cycle jumping technique such that, rather than applying a specified number of additional cycles, a specified local damage increment is specified and the number of additional cycles to reach this damage increment is calculated. In this way, the effect of stress redistribution due to damage-induced stiffness change is captured, but the fatigue simulations remain computationally efficient. The model is compared to experimental fatigue life data for two composite facesheet/foam core sandwich panels, demonstrating very good agreement.

Bednarcyk, Brett A.↗

Adaptive Controller Effects on Pilot Behavior

Adaptive control provides robustness and resilience for highly uncertain, and potentially unpredictable, flight dynamics characteristic. Some of the recent flight experiences of pilot-in-the-loop with an adaptive controller have exhibited unpredicted interactions. In retrospect, this is not surprising once it is realized that there are now two adaptive controllers interacting, the software adaptive control system and the pilot. An experiment was conducted to categorize these interactions on the pilot with an adaptive controller during control surface failures. One of the objectives of this experiment was to determine how the adaptation time of the controller affects pilots. The pitch and roll errors, and stick input increased for increasing adaptation time and during the segment when the adaptive controller was adapting. Not surprisingly, altitude, cross track and angle deviations, and vertical velocity also increase during the failure and then slowly return to pre-failure levels. Subjects may change their behavior even as an adaptive controller is adapting with additional stick inputs. Therefore, the adaptive controller should adapt as fast as possible to minimize flight track errors. This will minimize undesirable interactions between the pilot and the adaptive controller and maintain maneuvering precision.

Trujillo, Anna C.↗

Additive Manufacturing, Design, Testing, and Fabrication: A Full Engineering Experience at JSC

I worked on several projects this term. While most projects involved additive manufacturing, I was also involved with two design projects, two testing projects, and a fabrication project. The primary mentor for these was Richard Hagen. Secondary mentors were Hai Nguyen, Khadijah Shariff, and fabrication training from James Brown. Overall, my experience at JSC has been successful and what I have learned will continue to help me in my engineering education and profession long after I leave. My 3D printing projects ranged from less than a 1 cubic centimeter to about 1 cubic foot and involved several printers using different printing technologies. It was exciting to become familiar with printing technologies such as industrial grade FDM (Fused Deposition Modeling), the relatively new SLA (Stereolithography), and PolyJet. My primary duty with the FDM printers was to model parts that came in from various sources to print effectively and efficiently. Using methods my mentor taught me and the Stratasys Insight software, I was able to minimize imperfections, hasten build time, improve strength for specific forces (tensile, shear, etc...), and reduce likelihood of a print-failure. Also using FDM, I learned how to repair a part after it was printed. This is done by using a special kind of glue that chemically melts the two faces of plastic parts together to form a fused interface. My first goal with SLA technology was to bring the printer back to operational readiness. In becoming familiar with the Pegasus SLA printer, I researched the leveling, laser settings, and different vats to hold liquid material. With this research, I was successfully able to bring the Pegasus back online and have successfully printed multiple sample parts as well as functional parts. My experience with PolyJet technology has been focused on an understanding of the abilities/limits, costs, and the maintenance for daily use. Still upcoming will be experience with using a composite printer that uses FDM technology to print plastic while laying an internal filament of Kevlar or carbon-fiber inside the printed material. It has been incredible being exposed to this range of technologies and I feel very fortunate to be ready for virtually any kind of printing technology I come across in the future. Design work played a part in my internship this term as well. Working with Hai Nguyen, I was able to design a set of testing tips and a test frame for use with an Arcjet. The testing tips will be made of several different materials that will possibly be used in the heat shield of the Orion space craft. These designs included technical drawings that were presented to the fabrication shop. The frame design was created from 80/20 (a popular brand of frame construction equipment) and included an order form with pricing for fabrication. An independent design was also done for the virtual reality lab. This design was to create a hand-controller based on a previous design. This final design was sent directly to a 3D printer without technical drawings. Overall, my design work has given me experience with using 80/20, helped improve my CAD (Computer Aided Design) proficiency, and increased my knowledge of how to set up technical drawings for fabrication. The final role I have played in this internship has been to assist with testing of the inflatable technology materials working Khadijah Shariff. I began the internship assisting with permeability testing with the initial plan to continue the testing independently after training. Unfortunately, the testing apparatus suffered a technical failure and had funding pulled which cancelled that portion of the project. Further testing with inflatable technology continued with tensile testing of various stitching methods. This testing took place over a two-days and concluded successfully. Final testing was to be more tensile testing but of straps used to connect various inflatable sections. Unfortunately, the needed grips for the tensile tests could not be located and put the testing on hold. It is possible this round of testing will still take place by the end of the internship if the grips can be found. Overall, this portion of the internship has helped me become familiar with one kind of permeability test as well as a popular tensile/compression testing machine. Finally, I also had the chance to be trained using a metal lathe for making very small tips for a soldering iron. These tips will be used to melt threaded brass inserts into 3D printed plastic. Working with James Brown, I was able to successfully machine a brass rod down to as little as 0.064 inches plus or minus 0.001 inches. It was very rewarding to learn how to best use the machine and become familiar with a skill that will undoubtedly be used again in the future. I have been told by several professional engineers that learning to use a lathe and mill will be invaluable skills in the field. This knocks 50 percent of that goal off and I look forward to learning the mill at some point in the future. As is apparent with this list of projects, my internship was not focused on a single over-arching goal. Instead, I was able to gain experience in a myriad of very different areas. I feel like my time here was spent bouncing from one project to the next. Though sometimes difficult to switch gears, it was very rewarding to be a part of so much in so little time. My career and education will be positively impacted by what I have learned at JSC. My experience with 3D printing has improved my ability to handle many issues that may come up in the future with multiple different technologies. The design work I took part in, especially creating technical drawings, will help me better present designs to any engineer or shop I will encounter. My testing experience has helped me become familiar with a popular kind of tensile test machine that will likely be similar to the kinds I will encounter in the future. Finally, my experience with fabrication has given me a rare opportunity, as an engineer, to take part in the fabrication of a part. This experience will help me better tailor my future designs for the manufacturing process and has given me an appreciation for detailed/delicate machining work. My experience at JSC has been successful and will continue to assist me for a long time within the engineering field.

Zusack, Steven↗

Representing Matrix Cracks Through Decomposition of the Deformation Gradient Tensor in Continuum Damage Mechanics Methods

A method is presented to represent the large-deformation kinematics of intraply matrix cracks and delaminations in continuum damage mechanics (CDM) constitutive material models. The method involves the additive decomposition of the deformation gradient tensor into 'crack' and 'bulk material' components. The response of the intact bulk material is represented by a reduced deformation gradient tensor, and the opening of an embedded cohesive interface is represented by a normalized cohesive displacement-jump vector. The rotation of the embedded interface is tracked as the material deforms and as the crack opens. The distribution of the total local deformation between the bulk material and the cohesive interface components is determined by minimizing the difference between the cohesive stress and the bulk material stress projected onto the cohesive interface. The improvements to the accuracy of CDM models that incorporate the presented method over existing approaches are demonstrated for a single element subjected to simple shear deformation and for a finite element model of a unidirectional open-hole tension specimen. The material model is implemented as a VUMAT user subroutine for the Abaqus/Explicit finite element software. The presented deformation gradient decomposition method reduces the artificial load transfer across matrix cracks subjected to large shearing deformations, and avoids the spurious secondary failure modes that often occur in analyses based on conventional progressive damage models.

Leone, Frank A., Jr.↗

Improving Data Collection and Analysis Interface for the Data Acquisition Software of the Spin Laboratory at NASA Glenn Research Center

In jet engines, turbines spin at high rotational speeds. The forces generated from these high speeds make the rotating components of the turbines susceptible to developing cracks that can lead to major engine failures. The current inspection technologies only allow periodic examinations to check for cracks and other anomalies due to the requirements involved, which often necessitate entire engine disassembly. Also, many of these technologies cannot detect cracks that are below the surface or closed when the crack is at rest. Therefore, to overcome these limitations, efforts at NASA Glenn Research Center are underway to develop techniques and algorithms to detect cracks in rotating engine components. As a part of these activities, a high-precision spin laboratory is being utilized to expand and conduct highly specialized tests to develop methodologies that can assist in detecting predetermined cracks in a rotating turbine engine rotor. This paper discusses the various features involved in the ongoing testing at the spin laboratory and elaborates on its functionality and on the supporting data system tools needed to enable successfully running optimal tests and collecting accurate results. The data acquisition system and the associated software were updated and customized to adapt to the changes implemented on the test rig system and to accommodate the data produced by various sensor technologies. Discussion and presentation of these updates and the new attributes implemented are herein reported

Abdul-Aziz, Ali↗

Theoretical Development of an Orthotropic Elasto-Plastic Generalized Composite Material Model

The need for accurate material models to simulate the deformation, damage and failure of polymer matrix composites is becoming critical as these materials are gaining increased usage in the aerospace and automotive industries. While there are several composite material models currently available within LSDYNA (Livermore Software Technology Corporation), there are several features that have been identified that could improve the predictive capability of a composite model. To address these needs, a combined plasticity and damage model suitable for use with both solid and shell elements is being developed and is being implemented into LS-DYNA as MAT_213. A key feature of the improved material model is the use of tabulated stress-strain data in a variety of coordinate directions to fully define the stress-strain response of the material. To date, the model development efforts have focused on creating the plasticity portion of the model. The Tsai-Wu composite failure model has been generalized and extended to a strain-hardening based orthotropic yield function with a nonassociative flow rule. The coefficients of the yield function, and the stresses to be used in both the yield function and the flow rule, are computed based on the input stress-strain curves using the effective plastic strain as the tracking variable. The coefficients in the flow rule are computed based on the obtained stress-strain data. The developed material model is suitable for implementation within LS-DYNA for use in analyzing the nonlinear response of polymer composites.

Polymer Matrix Composites↗