Search NASA⌕ Search

SEARCH · Search NASA

Results for “Error Mitigation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 307 records · Page 17

Establishing Trust in NASA’s Artemis Campaign Computer-Human Interface (CHI) Implementation

The NASA Artemis program will return humans to the Moon. This time, with the help of commercial and international partners, the program's objective is a permanent moon base. The moon base infrastructure, including an orbiting moon station and moon surface assets, will be developed for astronauts to stay for the long haul to learn to live and work on another planet in preparation for an eventual Humans-to-Mars mission. As the roundtrip communication delays increase in deep space exploration, the crew will need more onboard systems autonomy and functionality to maintain and control the vehicle or habitat. These mission constraints will change the current Earth-based spacecraft to ground control support approach that will demand more safe, efficient, and effective Computer-Human Interface (CHI) control. For Artemis, CHI is defined as the elements that the crew interfaces with: audio, video, lighting, and crew controls subsystems. Understanding how CHI will need to evolve to support deep space missions will be critical for the Artemis program--especially crew controls, which is the focus of this paper. How does NASA ensure crew controls are reliable enough to control complex systems and prevent a catastrophic event due to human error--especially when the astronauts could be physiologically and/or psychologically impaired? NASA's approach to mitigating catastrophic hazards in human spaceflight system development such as crew controls, is through a holistic system engineering and Human System Integration methodology that focuses on incorporating NASA's Human-Rating Requirements-that ensures human performance characteristics to control/safely recover the crew from hazardous situations within the human interface design are considered. This paper discusses, at a high level, CHI for the Artemis program. Next, a discussion of what it means to human-rate a space system crew controls and how trust in the human-computer interface begins with the NASA human rating requirements. Finally, a discussion on how systems engineering and the human system integration process ensures that crew control implementation incorporates the NASA human-rating requirements.

artemis↗

First observation of RMP ELM mitigation on MAST Upgrade

Abstract The first experimental attempts at controlling edge localised modes (ELMs) via the application of resonant magnetic perturbations on the MAST Upgrade tokamak are reported. Using the linear MHD model MARS-F, the phase shift between the upper and lower coil rows ΔΦ was optimised for toroidal mode number n = 1 and n = 2 fields, to provide forward guidance to experiments. In low β N discharges, the application of n = 1 3D fields caused the ELM frequency f E L M to increase by over a factor 20 relative to the reference, and also induced a locked mode, which did not cause a plasma termination nor an H-L back transition. However when β N was raised, this induced locked mode caused plasma termination which precluded mitigation access. Initially, applying a numerically optimised n = 2 field had no effect. However applying a rigid toroidal shift to this field caused a locked mode disruption, demonstrating the presence of a substantial n = 2 error field. Coil current ramps were conducted with ΔΦ set at 6 different values, resulting in either locked mode disruptions or no effect, but mitigation with n = 2 fields was not established.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Modeling, Analyzing, and Mitigating Dissonance Between Alerting Systems

Alerting systems are becoming pervasive in process operations, which may result in the potential for dissonance or conflict in information from different alerting systems that suggests different threat levels and/or actions to resolve hazards. Little is currently available to help in predicting or solving the dissonance problem. This thesis presents a methodology to model and analyze dissonance between alerting systems, providing both a theoretical foundation for understanding dissonance and a practical basis from which specific problems can be addressed. A state-space representation of multiple alerting system operation is generalized that can be tailored across a variety of applications. Based on the representation, two major causes of dissonance are identified: logic differences and sensor error. Additionally, several possible types of dissonance are identified. A mathematical analysis method is developed to identify the conditions for dissonance originating from logic differences. A probabilistic analysis methodology is developed to estimate the probability of dissonance originating from sensor error, and to compare the relative contribution to dissonance of sensor error against the contribution from logic differences. A hybrid model, which describes the dynamic behavior of the process with multiple alerting systems, is developed to identify dangerous dissonance space, from which the process can lead to disaster. Methodologies to avoid or mitigate dissonance are outlined. Two examples are used to demonstrate the application of the methodology. First, a conceptual In-Trail Spacing example is presented. The methodology is applied to identify the conditions for possible dissonance, to identify relative contribution of logic difference and sensor error, and to identify dangerous dissonance space. Several proposed mitigation methods are demonstrated in this example. In the second example, the methodology is applied to address the dissonance problem between two air traffic alert and avoidance systems: the existing Traffic Alert and Collision Avoidance System (TCAS) vs. the proposed Airborne Conflict Management system (ACM). Conditions on ACM resolution maneuvers are identified to avoid dynamic dissonance between TCAS and ACM. Also included in this report is an Appendix written by Lee Winder about recent and continuing work on alerting systems design. The application of Markov Decision Process (MDP) theory to complex alerting problems is discussed and illustrated with an abstract example system.

Song, Lixia↗

Mitigating cosmic-ray-like correlated events with a modular quantum processor

Quantum processors based on superconducting qubits are being scaled to larger qubit numbers, enabling the implementation of small-scale quantum error-correction codes. However, catastrophic chip-scale correlated errors have been observed in these processors, attributed to, e.g., cosmic ray impacts, which challenge conventional error-correction codes such as the surface code. These events are characterized by a temporary but pronounced suppression of the qubit-energy relaxation times. Here, in this study, we explore the potential for modular quantum computing architectures to mitigate such correlated energy decay events. We measure cosmic-ray-like events in a quantum processor comprising a motherboard and two flip-chip bonded daughterboard modules, each module containing two superconducting qubits. We monitor the appearance of correlated qubit decay events within a single module and across the physically separated modules. We find that while decay events within one module are strongly correlated (over 85%), events in separate modules only display approximately 2% correlations. We also report coincident decay events in the motherboard and in either of the two daughterboard modules, providing further insight into the nature of these decay events. These results suggest that modular architectures, combined with bespoke errorcorrection codes, offer a promising approach for protecting future quantum processors from chip-scale correlated errors.

Wu, Xuntao [Univ. of Chicago, IL (United States)] ↗

Dual-frequency (Ka-band and G-band) radar estimates of liquid water content profiles in shallow clouds

The profile of the liquid water content (LWC) in clouds provides fundamental information for understanding the internal structure of clouds, their radiative effects, propensity to precipitate, and degree of entrainment and mixing with the surrounding environment. In principle, differential absorption techniques based on coincident dual-frequency radar reflectivity observations have the potential to provide the LWC profile. Previous differential frequency radar reflectivity (DFR) efforts were challenged by the fact that the measurable differential attenuation for small quantities of LWC is usually comparable to the system measurement error. This typically renders the retrieval impractical, as the uncertainty can become many times greater than the retrieved value itself. Theoretically, this drawback can be mitigated following two interconnected approaches: (1) increasing the frequency separation between the dual-frequency radar system to measure greater differential attenuation and (2) increasing the radar operating frequency to reduce the instrument measurement random error. Our recently developed 239 GHz radar was deployed during the Eastern Pacific Cloud Aerosol Precipitation Experiment (EPCAPE) along with a variety of collocated remote sensing and in situ instruments. We have combined Ka-band (35 GHz) and G-band (239 GHz) observations to retrieve the LWC from more than 100 vertical profiles of shallow clouds with typical amounts of LWC smaller than 1 g m -3 . We theoretically and experimentally demonstrate that the Ka-band and G-band pair of frequencies offers at least a 65 % relative improvement in the LWC retrieval sensitivity compared to previous works reported in the literature using lower-frequency radars. This new technique provides a missing capability to determine the LWC in the challenging low liquid water path (LWP) range (< 200 g m -2 ) and suggests a way forward to characterize microphysical and dynamical processes more precisely in shallow clouds.

54 ENVIRONMENTAL SCIENCES↗

The NASA Aviation Safety Program: Overview

In 1997, the United States set a national goal to reduce the fatal accident rate for aviation by 80% within ten years based on the recommendations by the Presidential Commission on Aviation Safety and Security. Achieving this goal will require the combined efforts of government, industry, and academia in the areas of technology research and development, implementation, and operations. To respond to the national goal, the National Aeronautics and Space Administration (NASA) has developed a program that will focus resources over a five year period on performing research and developing technologies that will enable improvements in many areas of aviation safety. The NASA Aviation Safety Program (AvSP) is organized into six research areas: Aviation System Modeling and Monitoring, System Wide Accident Prevention, Single Aircraft Accident Prevention, Weather Accident Prevention, Accident Mitigation, and Synthetic Vision. Specific project areas include Turbulence Detection and Mitigation, Aviation Weather Information, Weather Information Communications, Propulsion Systems Health Management, Control Upset Management, Human Error Modeling, Maintenance Human Factors, Fire Prevention, and Synthetic Vision Systems for Commercial, Business, and General Aviation aircraft. Research will be performed at all four NASA aeronautics centers and will be closely coordinated with Federal Aviation Administration (FAA) and other government agencies, industry, academia, as well as the aviation user community. This paper provides an overview of the NASA Aviation Safety Program goals, structure, and integration with the rest of the aviation community.

Shin, Jaiwon↗

Establishing Trust in NASA’s Artemis Program Computer-Human Interface (CHI) Implementation

The NASA Artemis program will return humans to the moon. This time, with the help of commercial and international partners, the program’s objective is a permanent moon base. The moon base infrastructure, including an orbiting moon station and moon surface assets, will be developed for astronauts to stay for the long haul to learn to live and work on another planet in preparation for an eventual Humans-to-Mars mission. As the roundtrip communication delays increase in deep space exploration, more onboard systems autonomy and functionality will be needed to maintain and control the vehicle or habitat. These mission constraints will change the current Earth-based spacecraft ground control support approach that will demand more safe, efficient, and effective Computer-Human Interface (CHI) control. For Artemis, CHI is defined as the elements that the crew interfaces with-audio, video, lighting, and crew controls. Understanding how CHI will need to evolve to support deep space missions will be critical for the Artemis program-especially crew controls which is the focus of this paper. How does NASA ensure crew controls are reliable to control complex systems and prevent a catastrophic event due to human error-especially when the astronauts could be physiologically and/or psychologically impaired? NASA’s approach to mitigating catastrophic hazards in human spaceflight system development such as crew controls is through a holistic system engineering and Human System Integration methodology that embraces NASA’s Human-Rating Requirements-ensuring human performance characteristics to control/safely recover the crew from hazardous situations within the human interface design are considered. This paper discusses, at a high level, CHI for the Artemis program. Next, a discussion of what it means to human-rate a space system crew controls and how trust in the human-computer interface begins with the NASA human rating requirements. Finally, a discussion on how systems engineering, and the human system integration process ensures that crew control implementation incorporates the NASA human-rating requirements.

Human-Rating↗

Enabling Pinpoint Landing (PPL) on Mars

Pinpoint landing (PPL) missions will deliver about 1000 kg of useful payload to the surface of Mars. Mid-to-high latitude landing site compatibility is sought which should provide the means to land at sites up to 2.5 km above Mars mean surface altitude. A dispersion and control analysis process is presented which helps to identify the effects of PPL error drivers, quantify the effect of dispersions on landing error and quantify the landing position control capability/authority along the entry path. An entry/descent/landing (EDL) profile is provided. Guided aeroshell is the baseline for all candidate Mars atmospheric entry architectures. A two-stage architecture is considered for the aerodynamic decelerator descent phase: supersonic parachute plus guided subsonic parachute or high-Mach inflatable decelerator plus guided subsonic parachute. The powered descent phase uses propulsive descent stage for soft landing and final error reduction maneuvers. Studies have found that the aeroshell entry face dispersions can be large, but closed-loop guidance can null out resulting errors to within about 2 km. Additionally, projected parachute control is inadequate to correct worst case dispersions without wind forecast data. To mitigate the problems dispersions due to atmospheric uncertainty can be reduced by providing on-board external means to measure density and winds ahead of the vehicle, higher L/D control authority options for the subsonic parachute phase can be investigated, and decelerators with control authority options for the supersonic descent phase can be examined. A navigation error analysis and wind effects summary are included.

aerodynamic↗

Potential effects of the introduction of the discrete address beacon system data link on air/ground information transfer problems

This study of Aviation Safety Reporting System reports suggests that benefits should accure from implementation of discrete address beacon system data link. The phase enhanced terminal information system service is expected to provide better terminal information than present systems by improving currency and accuracy. In the exchange of air traffic control messages, discrete address insures that only the intended recipient receives and acts on a specific message. Visual displays and printer copy of messages should mitigate many of the reported problems associated with voice communications. The problems that remain unaffected include error in addressing the intended recipient and messages whose content is wrong but are otherwise correct as to format and reasonableness.

Grayson, R. L.↗

Transmission Scheduling and Routing Algorithms for Delay Tolerant Networks

The challenges of data processing, transmission scheduling and routing within a space network present a multi-criteria optimization problem. Long delays, intermittent connectivity, asymmetric data rates and potentially high error rates make traditional networking approaches unsuitable. The delay tolerant networking architecture and protocols attempt to mitigate many of these issues, yet transmission scheduling is largely manually configured and routes are determined by a static contact routing graph. A high level of variability exists among the requirements and environmental characteristics of different missions, some of which may allow for the use of more opportunistic routing methods. In all cases, resource allocation and constraints must be balanced with the optimization of data throughput and quality of service. Much work has been done researching routing techniques for terrestrial-based challenged networks in an attempt to optimize contact opportunities and resource usage. This paper examines several popular methods to determine their potential applicability to space networks.

Space Networking↗

Error-controlled Progressive Retrieval of Scientific Data under Derivable Quantities of Interest

The unprecedented amount of scientific data has introduced heavy pressure on the current data storage and transmission systems. Progressive compression has been proposed to mitigate this problem, which offers data access with on-demand precision. However, existing approaches only consider precision control on primary data, leaving uncertainties on the quantities of interest (QoIs) derived from it. In this work, we present a progressive data retrieval framework with guaranteed error control on derivable QoIs. Our contributions are three-fold. (1) We carefully derive the theories to strictly control QoI errors during progressive retrieval. Our theory is generic and can be applied to any QoIs that can be composited by the basis of derivable QoIs proved in the paper. (2) We design and develop a generic progressive retrieval framework based on the proposed theories, and optimize it by exploring feasible progressive representations. (3) We evaluate our framework using five real-world datasets with a diverse set of QoIs. Experiments demonstrate that our framework can faithfully respect any user-specified QoI error bounds in the evaluated applications. This leads to over 2.02× performance gain in data transfer tasks compared to transferring the primary data while guaranteeing a QoI error that is less than 1E-5.

Wu, Xuan↗

Towards robust laser beam propagation in atmospheric turbulence

High-fidelity optical propagation through the atmosphere is essential for free-space optical technologies, including laser-based remote sensing and optical communication. However, atmospheric turbulence severely distorts beams and compromises system performance. In this work, we employ hypergeometric-Gaussian (HyGG) vortex beams as probes to characterize and mitigate atmospheric turbulence. Using over 250,000 experimental and simulated frames, we show that refining the power spectrum density (PSD) can reduce numerical prediction errors by up to 79.8%. Concurrently, experimental observations supported by numerical simulations demonstrate that HyGG beams exhibit superior turbulence resilience across multiple metrics compared to conventional Gaussian beams, particularly in their ability to withstand over 5 times stronger turbulence while maintaining similar intensity fluctuations. These dual investigations, on both turbulence mitigation and robust beam solutions, converge to form a unified strategy for enhancing free-space optical system performance. Collectively, our findings provide new insights into light–turbulence interactions and highlight the practical utility of vortex beams under atmospheric conditions.

Zhang, Boyu↗

Soft-Decision-Data Reshuffle to Mitigate Pulsed Radio Frequency Interference Impact on Low-Density-Parity-Check Code Performance

This presentation briefly discusses a research effort on mitigation techniques of pulsed radio frequency interference (RFI) on a Low-Density-Parity-Check (LDPC) code. This problem is of considerable interest in the context of providing reliable communications to the space vehicle which might suffer severe degradation due to pulsed RFI sources such as large radars. The LDPC code is one of modern forward-error-correction (FEC) codes which have the decoding performance to approach the Shannon Limit. The LDPC code studied here is the AR4JA (2048, 1024) code recommended by the Consultative Committee for Space Data Systems (CCSDS) and it has been chosen for some spacecraft design. Even though this code is designed as a powerful FEC code in the additive white Gaussian noise channel, simulation data and test results show that the performance of this LDPC decoder is severely degraded when exposed to the pulsed RFI specified in the spacecraft s transponder specifications. An analysis work (through modeling and simulation) has been conducted to evaluate the impact of the pulsed RFI and a few implemental techniques have been investigated to mitigate the pulsed RFI impact by reshuffling the soft-decision-data available at the input of the LDPC decoder. The simulation results show that the LDPC decoding performance of codeword error rate (CWER) under pulsed RFI can be improved up to four orders of magnitude through a simple soft-decision-data reshuffle scheme. This study reveals that an error floor of LDPC decoding performance appears around CWER=1E-4 when the proposed technique is applied to mitigate the pulsed RFI impact. The mechanism causing this error floor remains unknown, further investigation is necessary.

Ni, Jianjun David↗

Access and limits of RMP ELM suppression with n = 1 fields in DIII-D

This work reports on DIII-D experiments aimed at extending resonant magnetic perturbation (RMP) suppression of edge localized modes (ELMs) to n = 1 fields, where n is the toroidal mode number. Modeling of the 3D ideal MHD plasma response to the RMPs using the GPEC code is used to quantify edge and core resonant fluxes, guiding experimental strategies to increase plasma resilience against core error field penetration, optimize multicoil phasing, and explore higher q 95 operation. In DIII-D, ELM mitigation is regularly observed across a wide range of n = 1 RMP scenarios. A ∼100 ms phase of complete ELM suppression was achieved at q 95 ∼ 3.9 using an odd-parity coil configuration. The suppressed phase exhibited clear signatures of RMP ELM suppression, including the elimination of Dα spikes, increased pedestal rotation, enhanced magnetic response, and elevated broadband density turbulence. An optimized coil configuration for edge-to-core resonant flux did show increased edge resonance indicated by increased density pumpout, but did not yield RMP ELM suppression. At q 95 ∼ 5.1, a bifurcation to a grassy-like ELM regime occurred, while large type-I ELMs persisted. These results demonstrate progress in experimental access to n = 1 RMP ELM suppression in DIII-D, motivating further study for robust access. This work also highlights the potential role of 3D edge stability as well as rational surface alignment in RMP ELM suppression access, which has important implications for the use of low-n RMPs in future reactor-scale devices.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Spike-Free Adaptive Sliding Mode Control: Application to Permanent Magnet Synchronous Motors

A new methodology for adaptive sliding mode control (ASMC) has been widely used to improve the control performance in various systems. This method exhibits several advantages, including low sliding mode control (SMC) chattering, no knowledge of the system disturbance bound, and no overestimation of the control gain. Despite its advantages, this method can be hampered by the spike phenomenon, slow control gain convergence, and difficulty in achieving optimal performance under varying disturbances. Consequently, this article proposes a spike-free ASMC method with a disturbance observer (DOB) to address these problems. Previous ASMC methods have been analyzed via simulations to verify the aforementioned problems. Here, this analysis highlights the need for disturbance compensation and improvements in the SMC gain adaptation law. Therefore, a DOB is designed to mitigate the spike phenomenon by compensating for disturbances. Subsequently, an SMC gain adaptation law based on disturbance error estimation is designed to eliminate the spike phenomenon completely. The proposed adaptation law makes the SMC gain to converge to a slightly higher value than the disturbance estimation error. Consequently, the proposed method not only eliminates the spike phenomenon, but also ensures optimal performance under varying disturbances. The performance of the proposed method is experimen tally validated through a comparative study.

42 ENGINEERING↗

HRA Aerospace Challenges

Compared to equipment designed to perform the same function over and over, humans are just not as reliable. Computers and machines perform the same action in the same way repeatedly getting the same result, unless equipment fails or a human interferes. Humans who are supposed to perform the same actions repeatedly often perform them incorrectly due to a variety of issues including: stress, fatigue, illness, lack of training, distraction, acting at the wrong time, not acting when they should, not following procedures, misinterpreting information or inattention to detail. Why not use robots and automatic controls exclusively if human error is so common? In an emergency or off normal situation that the computer, robotic element, or automatic control system is not designed to respond to, the result is failure unless a human can intervene. The human in the loop may be more likely to cause an error, but is also more likely to catch the error and correct it. When it comes to unexpected situations, or performing multiple tasks outside the defined mission parameters, humans are the only viable alternative. Human Reliability Assessments (HRA) identifies ways to improve human performance and reliability and can lead to improvements in systems designed to interact with humans. Understanding the context of the situation that can lead to human errors, which include taking the wrong action, no action or making bad decisions provides additional information to mitigate risks. With improved human reliability comes reduced risk for the overall operation or project.

DeMott, Diana↗

ATTNChecker: Highly-Optimized Fault Tolerant Attention for Large Language Model Training

Large Language Models (LLMs) have demonstrated remarkable performance in various natural language processing tasks. However, the training of these models is computationally intensive and susceptible to faults, particularly in the attention mechanism, which is a critical component of transformer-based LLMs. In this paper, we investigate the impact of faults on LLM training, focusing on INF, NaN, and near-INF values in the computation results with systematic fault injection experiments. We observe the propagation patterns of these errors, which can trigger non-trainable states in the model and disrupt training, forcing the procedure to load from checkpoints. To mitigate the impact of these faults, we propose ATTNChecker, the first Algorithm-Based Fault Tolerance (ABFT) technique tailored for the attention mechanism in LLMs. ATTNChecker is designed based on fault propagation patterns of LLM and incorporates performance optimization to adapt to both system reliability and model vulnerability while providing lightweight protection for fast LLM training. Evaluations on four LLMs show that ATTNChecker on average incurs on average 7% overhead on training while detecting and correcting all extreme errors. Compared with the state-of-the-art checkpoint/restore approach, ATTNChecker reduces recovery overhead by up to 49×.

Liang, Yuhang [University of Alabama - Birmingham]↗

Mitigating Fading in Cislunar Communications: Application to the Human Landing System

NASA’s human exploration program is currently working towards landing astronauts on the surface of the Moon by 2024, close the lunar South Pole. To guarantee astronaut safety and maximize science data return, NASA is in the process of defining the communication architecture that will support all astronaut activities from launch to surface operations. Of particular interest to this paper are links from the lunar surface back to Earth without any intermediate relays. We show that the system geometry is such that antennas on the landing system will need to be pointed at low elevation angles, thus potentially causing multi-path fading effects not typically encountered in space communications. This paper is organized in three parts. First, we characterize the multi-path fading effects expected in links between the lunar South Pole and Earth and show that for moderate data rates (less than 1 Mbps) the links suffer from slow fading. We then show that for this operations regime the performance of forward error correction schemes is significantly worse for traditional Additive White Gaussian Noise channels. Finally, we investigate multi-copy mechanisms to mitigate the effects of fading, most notably repetition schemes and Automatic Repeat Request.

Sanchez Net, Marc↗