Search NASA⌕ Search

SEARCH · Search NASA

Results for “fault prevalence”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

A Field and Laboratory Study to Characterize Fault Prevalence in Residential Comfort Systems

This report describes a project whose goal was to determine the prevalence of residential comfort system faults. The focus is on air conditioners and heat pumps, and on key faults that can occur during installation and have significant impacts on performance: (1) incorrect refrigerant charge level (undercharge or overcharge); (2) indoor coil airflow rate; (3) liquid line restrictions; (4) non-condensable gas in the refrigerant; and (5) duct leakage. In addition to quantifying fault prevalence, we collected metadata that could be studied to determine whether there are correlations that suggest drivers of fault prevalence, such as regional variations in practice, economic factors, climatic impacts, and others.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Sensor cost-effectiveness analysis for data-driven fault detection and diagnostics in commercial buildings

Data-driven building fault detection and diagnostics (FDD) is heavily dependent on sensors. However, common sensors from Building Automation Systems are not optimized to maximize accuracy in FDD. Installing additional sensors that provide more detailed building system information is key to maximizing the performance of FDD solutions. Here in this paper, we present a sensor cost analysis workflow to quantify the economic implications of installing new sensors for FDD using the concept of sensor threshold marginal cost (STMC). STMC does not represent actual sensor cost. Rather, it represents a target cost based on the economic benefit that would be realized through improved FDD performance and one or more specified economic criteria. We calculate STMCs for multiple possible fault types and use fault prevalence information to aggregate STMCs into a single dollar value to determine the cost-effectiveness of a potential sensor investment. We conducted a case study using Oak Ridge National Laboratory's Flexible Research Platform (FRP) test facility as a reference. The case study demonstrates the feasibility of the analysis and highlights the key cost considerations in sensor selection for FDD. The results also indicate that identifying and installing the few key sensor(s) is critical to cost-effectively improve FDD performance.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Development of a Unified Taxonomy for HVAC System Faults

Detecting and diagnosing HVAC faults is critical for maintaining building operation performance, reducing energy waste, and ensuring indoor comfort. An increasing deployment of commercial fault detection and diagnostics (FDD) software tools in commercial buildings in the past decade has significantly increased buildings’ operational reliability and reduced energy consumption. A massive amount of data has been generated by the FDD software tools. However, efficiently utilizing FDD data for ‘big data’ analytics, algorithm improvement, and other data-driven applications is challenging because the format and naming conventions of those data are very customized, unstructured, and hard to interpret. This paper presents the development of a unified taxonomy for HVAC faults. A taxonomy is an orderly classification of HVAC faults according to their characteristics and causal relations. The taxonomy includes fault categorization, physical hierarchy, fault library, relation model, and naming/tagging scheme. The taxonomy employs both a physical hierarchy of HVAC equipment and a cause-effect relationship model to reveal the root causes of faults in HVAC systems. A structured and standardized vocabulary library is developed to increase data representability and interpretability. The developed fault taxonomy can be used for HVAC system ‘big data’ analytics such as HVAC system fault prevalence analysis or the development of an HVAC FDD software standard. A common type of HVAC equipment-packaged rooftop unit (RTU) is used as an example to demonstrate the application of the developed fault taxonomy. Two RTU FDD software tools are used to show that after mapping FDD data according to the taxonomy, the meta-analysis of the multiple FDD reports is possible and efficient.

96 KNOWLEDGE MANAGEMENT AND PRESERVATION↗

Automated fault detection and diagnosis of airflow and refrigerant charge faults in residential HVAC systems using IoT-enabled measurements

While automated fault detection and diagnosis (AFDD) in residential heating, ventilation, and air-conditioning (HVAC) using smart thermostat data is gaining increasing attention in recent times, it still requires in-depth investigation for market adoption, especially with real-life data. Furthermore, this paper proposes an Internet of Things (IoT) - based approach that adds a smart sensor to the smart thermostat data to carry out AFDD. The approach uses a model which predicts enthalpy change across the evaporator and compares the prediction to the measured enthalpy change. Deviations which exceed analytically determined thresholds then signal faults in the HVAC system. The faults detected are either installation related or degradation related. Experimental tests were carried out in four homes located in Norman, Oklahoma. From the tests, installation issues like indoor/outdoor mismatch were detected in two homes, while a 30% low charge and low indoor airflow rate were detected in one home. The results show that the proposed AFDD algorithm was able to successfully detect two prevalent faults, namely low indoor airflow and low refrigerant charge. Unlike most of the smart thermostat-based approaches, the proposed IoT-based approach can detect and diagnose both faults but only require one additional sensor which is provided by smart thermostat manufacturers.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Physics Based Modeling and Prognostics of Electrolytic Capacitors

This paper proposes first principles based modeling and prognostics approach for electrolytic capacitors. Electrolytic capacitors have become critical components in electronics systems in aeronautics and other domains. Degradations and faults in DC-DC converter unit propagates to the GPS and navigation subsystems and affects the overall solution. Capacitors and MOSFETs are the two major components, which cause degradations and failures in DC-DC converters. This type of capacitors are known for its low reliability and frequent breakdown on critical systems like power supplies of avionics equipment and electrical drivers of electromechanical actuators of control surfaces. Some of the more prevalent fault effects, such as a ripple voltage surge at the power supply output can cause glitches in the GPS position and velocity output, and this, in turn, if not corrected will propagate and distort the navigation solution. In this work, we study the effects of accelerated aging due to thermal stress on different sets of capacitors under different conditions. Our focus is on deriving first principles degradation models for thermal stress conditions. Data collected from simultaneous experiments are used to validate the desired models. Our overall goal is to derive accurate models of capacitor degradation, and use them to predict performance changes in DC-DC converters.

Kulkarni, Chetan↗

Atomic-Scale Behavior of Radiation-Resistant ZnO under High-Energy Electron Bombardment

Understanding the atomic structure and defect characteristics of ZnO thin films is crucial for optimizing their electronic properties and performance in advanced applications. Here, we investigate the atomic structure and defect characteristics of atomic layer deposition (ALD)-grown ZnO thin films by using aberration-corrected scanning transmission electron microscopy (STEM). Atomic-resolution imaging identifies prevalent stacking faults, dipole disorder, and various grain boundary types, which are believed to influence the electronic properties of ZnO. Additionally, real-time electron beam exposure experiments demonstrate structural transformations, including crystal growth and surface rearrangements. These findings provide insights into the growth mechanisms of ALD ZnO under high-energy electron irradiation conditions, an important finding for the use of polycrystalline ZnO wide bandgap semiconductors in space-like conditions. In conclusion, our results underscore the capability of STEM in directly visualizing and quantifying atomic-scale defects and beam-induced transformations in radiation-resistant ZnO.

Defects↗

Development and Validation of Home Comfort System for Total Performance Deficiency/Fault Detection and Optimal Comfort Control

In this project, we developed and tested a learning-based home thermal model that facilitates the operation of a model predictive control (MPC)-based optimization agent and an automated fault detection and diagnosis (AFDD) agent. The home thermal model was constructed using a two-node resistor-capacitor model. Moreover, two accompanying parameter identification methods were introduced, least-squares and optimization. Based on the home thermal model, the MPC-based optimization agent was developed to optimize residential HVAC operation. Using two FDD methods, the AFDD agent was constructed to detect and diagnose two prevalent residential AC faults, airflow reduction and refrigerant undercharge. The home thermal model, along with the MPC-based optimization agent and AFDD agent, were tested at the Norman Test House, Miami Test House, Pacific Northwest National Laboratory (PNNL) Test House A, and PNNL Test House B. Finally, they were also field tested in nine demonstration homes with real occupants.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Fracture and Stress Evolution on Europa: New Insights Into Fracture Interpretation and Ice Thickness Estimates Using Fracture Mechanics Analyses

The work completed during the funding period has provided many important insights into fracturing behavior in Europa's ice shell. It has been determined that fracturing through time is likely to have been controlled by the effects of nonsynchronous rotation stresses and that as much as 720 deg of said rotation may have occurred during the visible geologic history. It has been determined that there are at least two distinct styles of strike-slip faulting and that their mutual evolutionary styles are likely to have been different, with one involving a significant dilational component during shear motion. It has been determined that secondary fracturing in perturbed stress fields adjacent to older structures such as faults is a prevalent process on Europa. It has been determined that cycloidal ridges are likely to experience shear stresses along the existing segment portions as they propagate, which affects propagation direction and ultimately induces tailcracking at the segment tip than then initiates a new cycle of cycloid segment growth. Finally, it has been established that mechanical methods (e.g., flexure analysis) can be used to determine the elastic thickness of the ice shell, which, although probably only several km thick, is likely to be spatially variable, being thinner under bands but thicker under ridged plains terrain.

Kattenhorn, Simon↗

Resilience Design Patterns: A Structured Approach to Resilience at Extreme Scale (V.2.0)

Reliability is a serious concern for future extreme-scale high-performance computing (HPC) systems. Projections based on the current generation of HPC systems and technology roadmaps suggest the prevalence of very high fault rates in future systems. The errors resulting from these faults will propagate and generate various kinds of failures, which may result in outcomes ranging from result corruptions to catastrophic application crashes. Therefore, the resilience challenge for extreme-scale HPC systems requires coordination between various hardware and software technologies that are capable of handling a broad set of fault models at accelerated fault rates. Also, due to practical limits on power consumption in future HPC systems, they are likely to embrace innovative architectures, increasing the levels of hardware and software complexities. Therefore, the techniques that seek to improve resilience must navigate the complex trade-off space between resilience and the overheads to power consumption and performance. While the HPC community has developed various resilience solutions, application-level techniques as well as system-based solutions, the solution space of HPC resilience techniques remains fragmented. There are no formal methods to integrate the various HPC resilience techniques into composite solutions, nor are there methods to holistically evaluate the adequacy and efficacy of such solutions in terms of their protection coverage, and their performance & power efficiency characteristics. Additionally, few implementations of current resilience solutions are portable to newer architectures and software environments that will be deployed on future systems. We developed a new structured approach to the management of HPC resilience using the concept of resilience-based design patterns. In general, a design pattern is a repeatable solution to a commonly occurring problem. We identified the well-known solutions that are commonly used to deal with faults, errors and failures in HPC systems. In the initial design patterns specification (version 1.0), we described the various solutions, which address specific problems in the design of resilient HPC environments, in the form of patterns. Each pattern describes a problem caused by a fault, error or failure event in an HPC environment, and then describes the core of the solution of the problem in such a way that this solution may be adapted to different systems and implemented at different layers of the system stack. The catalog of these resilience design patterns provides designers with a collection of design elements. To construct complete resilience solutions using combinations of various patterns, we defined a framework that enhances HPC designers' understanding of the important constraints and the opportunities for the design patterns to be implemented and deployed at various layers of the system stack. The design framework is also useful for establishing interfaces and mechanisms to coordinate flexible fault management across hardware and software components, as well as to consider the trade-off between performance, resilience, and power consumption when constructing a solution. The resilience design patterns specification version 1.1 included more detailed explanations of the pattern solutions, the context in which the patterns are applicable, and the implications for hardware or software design. It also provided several additional examples and detailed case studies to demonstrate the use of patterns to build realistic solutions. In version 1.2 of the specification document, we have improved the pattern descriptions, including graphical representations of the pattern components. These improvements are largely based on critical comments, feedback and suggestions received from pattern experts and readers of the previous versions of the specification. The pattern classification has been modified to further clarify the relationships between pattern categories. This version of the specification also introduces a pattern language for resilience design patterns. The pattern language presents the patterns in the catalog as a network, revealing the relations among the resilience patterns. The language provides designers with the means to explore alternative techniques for handling a specific fault model that may have different efficiency and complexity characteristics. Using the pattern language also enables the design and implementation of comprehensive resilience solutions as a set of interconnected resilience patterns that can be instantiated across layers of the system stack. The overall goal of this work is to provide hardware and software designers, as well as the users and operators of HPC systems, a systematic methodology for the design and evaluation of resilience technologies in HPC systems that keep scientific applications running to a correct solution in a timely and cost-efficient manner despite frequent faults, errors, and failures of various types. Version 2.0 expands the resilience design pattern classification and catalog to include self-stabilization patterns and reliability, availability and performance models for each structural pattern.

97 MATHEMATICS AND COMPUTING↗

Use of Very High-Resolution Optical Data for Landslide Mapping and Susceptibility Analysis Along the Karnali Highway, Nepal

The Karnali highway is a vital transport link and the only primary roadway that connects the remote Karnali region to the lowlands in Mid-Western Nepal. Every year there are reports of landslides blocking the road, making this area largely inaccessible. However, little effort has focused on systematically identifying landslides and landslide-prone areas along this highway. In this study, landslides were mapped with an object-based approach from very high-resolution optical satellite imagery obtained by the DigitalGlobe constellation in 2012 and PlanetScope in 2018. Landslides ranging from 10 to 30,496 sq.m were detected within a 3 km buffer along the highway. Most of the landslides were located at lower elevations (between 500–1500 m) and on steep south-facing slopes. Landslides tended to cluster closer to the highway, near drainage channels and away from faults. Landslides were also most prevalent within the Kuncha Formation geologic class, and the forested and agricultural land cover classes. A susceptibility map was then created using a logistic regression methodology to highlight patterns in landslide activity. The landslide susceptibility map showed a good prediction rate with an area under the curve (AUC) of 0.90. A total of 33% of the study arealies in high/very high susceptibility zones. The map highlighted the lower elevated areas between Bangesimal and Manma towns with the Kuncha Formation geologic class as being the most hazardous. The banks of the Karnali River, its tributaries and areas near the highway were also highly susceptible to landslides. The results highlight the potential of very high-resolution optical imagery for documenting detailed spatial information on landslide occurrence, which enables susceptibility assessment in remote and data scarce regions such as the Karnali highway.

Pukar Amatya↗

Barriers to Broader Utilization of Fault Detection Technologies for Improving Residential HVAC Equipment Efficiency

Faults in residential heating, ventilating, and air conditioning (HVAC) equipment may occur due to poor installation practices or develop over time, and these faults can negatively impact system efficiency, thermal comfort, and equipment lifespan. Automated fault detection and diagnostic (AFDD) technologies identify energy wasting HVAC faults, such as low indoor airflow and improper refrigerant charge, and guide technicians in improving system efficiency. For residential HVAC, AFDD consists of a range of fault detecting and diagnostic capabilities, sensor configurations, and target applications. AFDD technology can either be permanently installed by the original equipment manufacturer (OEM) using embedded sensors or as an add-on product either during or after installation. Additionally, several advanced installation tools and refrigerant gauge sets include AFDD features for temporary use during equipment installation and tune-ups. Some technologies can detect a fault but have limited diagnostic capabilities. For example, a single-point measurement from the home's thermostat or energy monitor can provide certain fault detection capability by analyzing the equipment runtime or energy consumption. These technologies, though limited at determining the cause of a given fault, may have significant energy savings potential due to their low cost and prevalence in the residential HVAC market. Despite the potential benefits, fault detection technologies face many technical and market barriers preventing broad adoption. Beyond the cost barriers due to the added sensor requirements and technology development, fault detection technologies face many implementation and adoption barriers such as installer training, customer awareness, standardized communication protocols, and methods of test for evaluating accuracy. The purpose of this whitepaper is to characterize market and technical barriers impeding broader utilization of fault detection technology for residential HVAC energy efficiency applications.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Modeling Air Handling Units to Create a Diverse Fault Dataset for FDD Innovation: Lessons Learned and Recommendations

As energy management and information systems (e.g., automated fault detection and diagnostics [AFDD] tools) become more prevalent in the commercial building stock, it is important to determine the effectiveness of these technologies by benchmarking their performance. The authors have been working to develop the largest publicly available dataset of HVAC fault datasets for performance benchmarking applications, covering the most common HVAC systems and designs including chiller plants, rooftop packaged units, dual duct air handling unit and single duct air handling units. This study covers the development, modeling, and validation of a synthetic fault dataset for the air handling unit (AHU), one of the most common HVAC configurations found in the commercial building stock. Despite this being a common system, real-world time series data are scarce and usually do not span a wide range of weather conditions. Due to this limitation, two detailed AHU models, which included the single duct AHU and dual duct AHU developed in the Modelica language and HVACSIM+ were employed to carry out annual simulations of numerous common sensor faults, mechanical faults, and control sequence faults. The fault inclusive data were then validated by comparing fault effects on system performance to expected symptoms. We summarize the nature of each fault and their impacts under different weather and operation conditions. We report some lessons learnt during the efforts of validating the high volumes of the FDD data sets. Finally, we highlight considerations for FDD developers that may want to use this dataset to assess their algorithms’ performance and their improvement over time.

Casillas, Armando↗

Development of a Annual Air Handling Unit Fault Dataset for FDD Tools: Lessons Learned and Considerations for FDD Developers

As energy management and information systems (e.g., automated fault detection and diagnostics [AFDD] tools) become more prevalent in the commercial building stock, it is important to determine the effectiveness of these technologies by benchmarking their performance. The authors have been working to develop the largest publicly available dataset of HVAC fault data for performance benchmarking applications, covering the most common HVAC systems and designs including chiller plants, rooftop packaged units, dual duct air handling units and single duct air handling units. This study covers the development, modeling, and validation of a synthetic fault dataset for a single duct air handling unit (AHU), one of the most common HVAC configurations found in the commercial building stock. Despite this being a common system, real-world time series data are scarce and usually do not span a wide range of weather conditions. Due to this limitation, a detailed AHU model was employed to carry out annual simulations of numerous common sensor and mechanical faults, which were then validated by comparing their effects on system performance to expected symptoms. We summarize the nature of each fault and their impacts under different weather and operation conditions. Finally, we highlight considerations for FDD developers that may want to use this dataset to assess their algorithms’ performance and their improvement over time.

Casillas, Armando↗

Digital system bus integrity

This report summarizes and describes the results of a study of current or emerging multiplex data buses as applicable to digital flight systems, particularly with regard to civil aircraft. Technology for pre-1995 and post-1995 timeframes has been delineated and critiqued relative to the requirements envisioned for those periods. The primary emphasis has been an assured airworthiness of the more prevalent type buses, with attention to attributes such as fault tolerance, environmental susceptibility, and problems under continuing investigation. Additionally, the capacity to certify systems relying on such buses has been addressed.

Eldredge, Donald↗

DER Inverter Control Fault Ride Through Model in Accordance with IEEE 1547-2018 Std

Distributed Energy Resources (DER) with smart inverters are becoming more prevalent as the need for renewable energy and grid stability increases. An important challenge arises when considering that inverterbased generation methods contribute less current during faults, rendering traditional overcurrent protection unsatisfactory. DERs have fault ride-through requirements when operating in high or low voltage, outlined by IEEE Std. 1547-2018. Faults cause the voltage to reach abnormal steady state magnitudes, depending on the fault resistance and fault type. There are several high voltage and low voltage ride-through zones defined by IEEE Std. 1547-2018. Each zone’s ride through duration decreases as the applicable voltage measurement, i.e., the phase RMS voltage, deviates from its nominal value. This presentation demonstrates the implementation of IEEE Std. 1547-2018 high and low voltage ridethrough grid support functions using a preexisting RSCAD model, discussing the challenges presented during this process. The implemented controls monitor the filtered phase voltages to have a more accurate reading of the applicable voltages. The controls sense the duration that the applicable voltage remains in a specific zone. The breaker trips and ceases energization to the grid when the duration is exceeded. The standard allows the operator to adjust the ride-through times and voltage zones from the default settings. These ranges are implemented into the runtime, which acts as the operator’s SCADA. The results show the accuracy of the voltage measurements, which remain within the IEEE Std. 1547-2018 for all cases.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Evolutionary Based Techniques for Fault Tolerant Field Programmable Gate Arrays

The use of SRAM-based Field Programmable Gate Arrays (FPGAs) is becoming more and more prevalent in space applications. Commercial-grade FPGAs are potentially susceptible to permanently debilitating Single-Event Latchups (SELs). Repair methods based on Evolutionary Algorithms may be applied to FPGA circuits to enable successful fault recovery. This paper presents the experimental results of applying such methods to repair four commonly used circuits (quadrature decoder, 3-by-3-bit multiplier, 3-by-3-bit adder, 440-7 decoder) into which a number of simulated faults have been introduced. The results suggest that evolutionary repair techniques can improve the process of fault recovery when used instead of or as a supplement to Triple Modular Redundancy (TMR), which is currently the predominant method for mitigating FPGA faults.

Larchev, Gregory V.↗

BRAINSTACK – A Platform for Artificial Intelligence & Machine Learning Collaborative Experiments on a Nano-Satellite

As space missions continue to become more ambitious, complex, and distant to Earth, the need for advanced on-board intelligent decision making to guide everything from mission operations to fault detection and recovery has become a major front of space research. While the prevalence of research on such Artificial Intelligence / Machine Learning (AI/ML) modules has exploded, the capacity to experimentally validate such modules in space in a rapid and inexpensive format has not. To this end, the Nano Orbital Workshop (NOW) group at NASA Ames Research Center has been at the forefront of performing initial flight evaluation tests of ‘commercially’ available AI/ML computational platforms via the TechEdSat (TES-n) flight series as part of what is programmatically referred to as the BRAINSTACK. BRAINSTACK will provide an orbital AI/ML evaluation laboratory where computational experiments are pre-loaded into memory prior to launch, and then executed as desired during the mission, with results reported back and program tweaks or new data sets uploaded as needed. Processors selected as part of the BRAINSTACK are of ideal size, packaging, and power consumption for easy integration into a cube satellite structure. These experiments have included the evaluation of small, high-performance GPUs and more recently, neuromorphic processors in LEO operations. Neuromorphic processors are of particular interest due to their superior computational power efficiency over GPUs. The first TES-n flight test of an Intel first-generation Loihi neuromorphic processor launched on January 13, 2022 and continues to operate in orbit despite almost no space environment modifications. The Intel Loihi Gen-1 is characterized by a 14nm 128-core Spiking Neural Network (SNN) able to support on-chip training. This experiment utilized a Loihi packaged in the ‘Kapoho Bay’ USB module, providing a relatively straight-forward interface to the bus avionics system. The Kapoho Bay was in turn managed by a host Intel Pentium single-board computer to handle scheduling of the AI/ML application payloads, and communications with the satellite vehicle manager. The recently released Intel Loihi Gen-2, able to support integer-valued spike payloads and produced using 7nm process, will form part of the basis of the evolving BRAINSTACK in the upcoming three TES-n/NOW flights. Additionally, it is planned to measure the radiation environment these processors experience to understand any degradation or computational artifacts caused by long term space radiation exposure on these novel architectures. This evolving flexible and collaborative environment involving various research teams across NASA and other organizations is intended to be a convenient orbital test platform from which many anticipated future space AI/ML applications may be initially tested.

Artificial Intelligence↗