Search NASA⌕ Search

SEARCH · Search NASA

Results for “Adaptive learning”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 145 records · Page 8

Modification of Eccentric Gaze-Holding

Clear vision and accurate localization of objects in the environment are prerequisites for reliable performance of motor tasks. Space flight confronts the crewmember with a stimulus rearrangement that requires adaptation to function effectively with the new requirements of altered spatial orientation and motor coordination. Adaptation and motor learning driven by the effects of cerebellar disorders may share some of the same demands that face our astronauts. One measure of spatial localization shared by the astronauts and those suffering from cerebellar disorders that is easily quantified, and for which a neurobiological substrate has been identified, is the control of the angle of gaze (the "line of sight"). The disturbances of gaze control that have been documented to occur in astronauts and cosmonauts, both in-flight and postflight, can be directly related to changes in the extrinsic gravitational environment and intrinsic proprioceptive mechanisms thus, lending themselves to description by simple non-linear statistical models. Because of the necessity of developing robust normal response populations and normative populations against which abnormal responses can be evaluated, the basic models can be formulated using normal, non-astronaut test subjects and subsequently extended using centrifugation techniques to alter the gravitational and proprioceptive environment of these subjects. Further tests and extensions of the models can be made by studying abnormalities of gaze control in patients with cerebellar disease. A series of investigations were conducted in which a total of 62 subjects were tested to: (1) Define eccentric gaze-holding parameters in a normative population, and (2) explore the effects of linear acceleration on gaze-holding parameters. For these studies gaze-holding was evaluated with the subjects seated upright (the normative values), rolled 45 degrees to both the left and right, or pitched back 30 and 90 degrees. In a separate study the further effects of acceleration on gaze stability was examined during centrifugation (+2 G (sub x) and +2 G (sub z) using a total of 23 subjects. In all of our investigations eccentric gaze-holding was established by having the subjects acquire an eccentric target (+/-30 degrees horizontal, +/- 15 degrees vertical) that was flashed for 750 msec in an otherwise dark room. Subjects were instructed to hold gaze on the remembered position of the flashed target for 20 sec. Immediately following the 20 sec period, subjects were cued to return to the remembered center position and to hold gaze there for an additional 20 sec. Following this 20 sec period the center target was briefly flashed and the subject made any corrective eye movement back to the true center position. Conventionally, the ability to hold eccentric gaze is estimated by fitting the natural log of centripetal eye drifts by linear regression and calculating the time constant (G) of these slow phases of "gaze-evoked nystagmus". However, because our normative subjects sometimes showed essentially no drift (tau (sub c) = m), statistical estimation and inference on the effect of target direction was performed on values of the decay constant theta = 1/(tau (sub c)) which we found was well modeled by a gamma distribution. Subjects showed substantial variance of their eye drifts, which were centrifugal in approximately 20 % of cases, and > 40% for down gaze. Using the ensuing estimated gamma distributions, we were able to conclude that rightward and leftward gaze holding were not significantly different, but that upward gaze holding was significantly worse than downward (p<0.05). We also concluded that vertical gaze holding was significantly worse than horizontal (p<0.05). In the case of left and right roll, we found that both had a similar improvement to horizontal gaze holding (p<0.05), but didn't have a significant effect on vertical gaze holding. For pitch tilts, both tilt angles significantly decreased gaze-holding ility in all directions (p<0.05). Finally, we found that hyper-g centrifugation significantly decreased gaze holding ability in the vertical plane. The main findings of this study are as follows: (1) vertical gaze-holding is less stable than horizontal, (2) gaze-holding to upward targets is less stable than to downward targets, (3) tilt affects gaze holding, and (4) hyper-g affects gaze holding. This difference between horizontal and vertical gaze-holding may be ascribed to separate components of the velocity-to-position neural integrator for eye movements, and to differences in orbital mechanics. The differences between upward and downward gaze-holding may be ascribed to an inherent vertical imbalance in the vestibular system. Because whole body tilt and hyper-g affects gaze-holding, it is implied that the otolith organs have direct connections to the neural integrator and further studies of astronaut gaze-holding are warranted. Our statistical method for representing the range of normal eccentric gaze stability can be readily applied to normals who maybe exposed to environments which may modify the central integrator and require monitoring, and to evaluate patients with gaze-evoked nystagmus by comparing to the above established normative criteria.

Reschke, M. F.↗

Applying Fortran 90 and Object-Oriented Techniques to Scientific Applications

High-performance parallel computing is having a profound impact on the size and complexity of physical problems which can be modeled. This impact has been a long time in coming, because the learning curve in adapting to this new world of computing is steeper than was imagined.

Fortran 90 parallel computing C++↗

Forward Contamination of Ocean Worlds: A Stakeholder Conversation

A fundamental requirement for space missions designed to touch “potential habitats” is the single number 10−4, the allowable probability of a single Earth organism contaminating the potential habitat. Many aspects of a mission that affect its complexity and cost – hardware design and manufacture, assembly and test, and mission operations – are driven by this value, so it is important, on the threshold of an era of exploring ocean worlds, to have confidence in it. Yet despite its long pedigree and occasional reviews, we find that the current requirement lacks programmatically defensible justification. At issue are three weaknesses: 1) microbial biology, in particular the science of extremophiles, is a rapidly changing field; 2) forward contamination is both a scientific and an ethical issue, yet no ethics-based conversation is apparent within policy-setting circles; 3) because of these two factors, policy-setting cannot be static. We review the history of the requirement; how the evolving understanding of biology could drive it up or down; how the forward-contamination hazard relates to risk-management practice and to the ethics profession; and how a contemporary stakeholder conversation could adapt lessons already learned by other fields.

Waltemathe, Michael↗

Assessment of Postflight Locomotor Performance Utilizing a Test of Functional Mobility: Strategic and Adaptive Responses

Space flight induces adaptive modification in sensorimotor function, allowing crewmembers to operate in the unique microgravity environment. This adaptive state, however, is inappropriate for a terrestrial environment. During a re-adaptation period upon their return to Earth, crewmembers experience alterations in sensorimotor function, causing various disturbances in perception, spatial orientation, posture, gait, and eye-head coordination. Following long duration space flight, sensorimotor dysfunction would prevent or extend the time required to make an emergency egress from the vehicle; compromising crew safety and mission objectives. We are investigating two types of motor learning that may interact with each other and influence a crewmember's ability to re-adapt to Earth's gravity environment. In strategic learning, crewmembers make rapid modifications in their motor control strategy emphasizing error reduction. This type of learning may be critical during the first minutes and hours after landing. In adaptive learning, long-term plastic transformations occur, involving morphological changes and synaptic modification. In recent literature these two behavioral components have been associated with separate brain structures that control the execution of motor strategies: the strategic component was linked to the posterior parietal cortex and the adaptive component was linked to the cerebellum (Pisella, et al. 2004). The goal of this paper was to demonstrate the relative contributions of the strategic and adaptive components to the re-adaptation process in locomotor control after long duration space flight missions on the International Space Station (ISS). The Functional Mobility Test (FMT) was developed to assess crewmember s ability to ambulate postflight from an operational and functional perspective. Sixteen crewmembers were tested preflight (3 sessions) and postflight (days 1, 2, 4, 7, 25) following a long duration space flight (approx 6 months) on the ISS. We have further analyzed the FMT data to characterize strategic and adaptive components during the postflight readaptation period. Crewmembers walked at a preferred pace through an obstacle course set up on a base of 10 cm thick medium density foam (Sunmate Foam, Dynamic Systems, Inc., Leicester, NC). The 6.0m X 4.0m course consisted of several pylons made of foam; a Styrofoam barrier 46.0cm high that crewmembers stepped over; and a portal constructed of two Styrofoam blocks, each 31cm high, with a horizontal bar covered by foam and suspended from the ceiling which was adjusted to the height of the crewmember s shoulder. The portal required crewmembers to bend at the waist and step over a barrier simultaneously. All obstacles were lightweight, soft and easily knocked over. Crewmembers were instructed to walk through the course as quickly and as safely as possible without touching any of the objects on the course. This task was performed three times in the clockwise direction and three times in the counterclockwise direction that was randomly chosen. The dependent measures for each trial were: time to complete the course (seconds) and the number of obstacles touched or knocked down. For each crewmember, the time to complete each FMT trial from postflight days 1, 2, 4, 7 and 25 were further analyzed. A single logarithmic curve using a least squares calculation was fit through these data to produce a single comprehensive curve (macro). This macro curve composed of data spanning 25 days, illustrates the re-adaptive learning function over the longer time scale term. Additionally, logarithmic curves were fit to the 6 data trials within each individual post flight test day to produce 5 separate daily curves. These micro curves, produced from data obtained over the course of minutes, illustrates the strategic learning function exhibited over a relative shorter time scale. The macro curve for all subjects exhibited adaptive motor learning patterns over the 25 day period. Howev, 9/16 crewmembers exhibited significant strategic motor learning patterns in their micro curves, as defined by m > 1 in the equation of the line y=m*LN(x) +b. These data indicate that postflight recovery in locomotor function involves both strategic and adaptive mechanisms. Future countermeasures will be designed to enhance both recovery processes.

Warren, L. E.↗

The NASA F-15 Intelligent Flight Control Systems: Generation II

The Second Generation (Gen II) control system for the F-15 Intelligent Flight Control System (IFCS) program implements direct adaptive neural networks to demonstrate robust tolerance to faults and failures. The direct adaptive tracking controller integrates learning neural networks (NNs) with a dynamic inversion control law. The term direct adaptive is used because the error between the reference model and the aircraft response is being compensated or directly adapted to minimize error without regard to knowing the cause of the error. No parameter estimation is needed for this direct adaptive control system. In the Gen II design, the feedback errors are regulated with a proportional-plus-integral (PI) compensator. This basic compensator is augmented with an online NN that changes the system gains via an error-based adaptation law to improve aircraft performance at all times, including normal flight, system failures, mispredicted behavior, or changes in behavior resulting from damage.

Buschbacher, Mark↗

Controlling a truck with an adaptive critic temporal difference CMAC design

In this study, CMAC (Cerebellar Model Articulated Controller) neural architectures are shown to be viable for the purposes of real-time learning and control. An adaptive critic temporal difference neurocontrol design has been implemented that learns in real-time how to back up a trailer truck along a fixed straight line trajectory. The truck backer-upper experiment is a standard performance measure in the neural network literature, but previously the training of the controllers was done off-line. With the CMAC neural architectures, it was possible to train the neurocontrollers on-line in real-time on a MS-DOS PC 386.

Shelton, Robert O.↗

Gravity Sensor Plasticity in the Space Environment

The ability of the brain to learn from experience and to adapt to new environments is recognized to be profound. This ability, called 'neural plasticity,' depends directly on properties of neurons (nerve cells) that permit them to change in dimension, sprout new parts called spines, change the shape and/or size of existing parts, and to generate, alter, or delete synapses. (Synapses are communication sites between neurons.) These neuronal properties are most evident during development, when evolution guides the laying down of a general plan of the nervous system. However, once a nervous system is established, experience interacts with cellular and genetic mechanisms and the internal milieu to produce unique neuronal substrates that define each individual. The capacity for experience-related neuronal growth in the brain, as measured by the potential for synaptogenesis, is speculated to be in the trillions of synapses, but the range of increment possible for any one part of the nervous system is unknown. The question has been whether more primitive endorgans such as gravity sensors of the inner ear have a capacity for adaptive change, since this is a form of learning from experience.

Ross, Muriel D.↗

Completing and Adapting Models of Biological Processes

We present a learning-based method for model completion and adaptation, which is based on the combination of two approaches: 1) R2D2C, a technique for mechanically transforming system requirements via provably equivalent models to running code, and 2) automata learning-based model extrapolation. The intended impact of this new combination is to make model completion and adaptation accessible to experts of the field, like biologists or engineers. The principle is briefly illustrated by generating models of biological procedures concerning gene activities in the production of proteins, although the main application is going to concern autonomic systems for space exploration.

Margaria, Tiziana↗

Nonlinear functional approximation with networks using adaptive neurons

A novel mathematical framework for the rapid learning of nonlinear mappings and topological transformations is presented. It is based on allowing the neuron's parameters to adapt as a function of learning. This fully recurrent adaptive neuron model (ANM) has been successfully applied to complex nonlinear function approximation problems such as the highly degenerate inverse kinematics problem in robotics.

Tawel, Raoul↗

Closing the Certification Gaps in Adaptive Flight Control Software

Over the last five decades, extensive research has been performed to design and develop adaptive control systems for aerospace systems and other applications where the capability to change controller behavior at different operating conditions is highly desirable. Although adaptive flight control has been partially implemented through the use of gain-scheduled control, truly adaptive control systems using learning algorithms and on-line system identification methods have not seen commercial deployment. The reason is that the certification process for adaptive flight control software for use in national air space has not yet been decided. The purpose of this paper is to examine the gaps between the state-of-the-art methodologies used to certify conventional (i.e., non-adaptive) flight control system software and what will likely to be needed to satisfy FAA airworthiness requirements. These gaps include the lack of a certification plan or process guide, the need to develop verification and validation tools and methodologies to analyze adaptive controller stability and convergence, as well as the development of metrics to evaluate adaptive controller performance at off-nominal flight conditions. This paper presents the major certification gap areas, a description of the current state of the verification methodologies, and what further research efforts will likely be needed to close the gaps remaining in current certification practices. It is envisioned that closing the gap will require certain advances in simulation methods, comprehensive methods to determine learning algorithm stability and convergence rates, the development of performance metrics for adaptive controllers, the application of formal software assurance methods, the application of on-line software monitoring tools for adaptive controller health assessment, and the development of a certification case for adaptive system safety of flight.

Jacklin, Stephen A.↗

Development of Machine Learning-Derived Microbiological and Immune Signatures: Applications in Adaptive Risk Assessment of Infectious Disease During Spaceflight

Infectious diseases represent an urgent risk for spaceflight with consequences ranging from loss in crew performance to crew incapacitation or loss of life should an outbreak occur. The resident environmental microbiome on the International Space Station has been monitored through routine surveillance over almost twenty years, beginning with culture-based microbial detection which has advanced to molecular methods in recent years. This has created a wealth of data that we have begun mining to define the microbial ecology of the ISS. Summarized here is our analysis of data from the historical microbial population defined by culture-based monitoring from the past two decades, organized by their likelihood to cause disease into clinical categories. As expected, many residents of the normal microflora in environments where people work and live were detected. However, some known pathogens were also detected. As the spaceflight environment can predispose humans to infection, crew health records were used to source additional data for the set to uncover clinical relevance. Data mining was performed on crew health records to capture adverse health events that may be related to infectious disease. Machine learning, specifically Random Forest analysis, was used to analyze the microbial and crew health datasets. The symptom categories were not explained by the ranked bacteria, due to lack of sufficient data for some categories and due to poor ranking of the pathogens for others. Poor ranking of the bacteria could be due to the clinical symptoms being linked to other disease-causing factors, such as allergy or viral infection. These findings suggest a lack of relationship between bacteria detected on surfaces in the ISS and historical health events experienced by astronauts.

Kristyn Hoffman↗

Neuromorphic learning of continuous-valued mappings in the presence of noise: Application to real-time adaptive control

The ability of feed-forward neural net architectures to learn continuous-valued mappings in the presence of noise is demonstrated in relation to parameter identification and real-time adaptive control applications. Factors and parameters influencing the learning performance of such nets in the presence of noise are identified. Their effects are discussed through a computer simulation of the Back-Error-Propagation algorithm by taking the example of the cart-pole system controlled by a nonlinear control law. Adequate sampling of the state space is found to be essential for canceling the effect of the statistical fluctuations and allowing learning to take place.

Troudet, Terry↗

Adaptive Stress Testing: Finding Likely Failure Events with Reinforcement Learning

Finding the most likely path to a set of failure states is important to the analysis of safety-critical systems that operate over a sequence of time steps, such as aircraft collision avoidance systems and autonomous cars. In many applications such as autonomous driving, failures cannot be completely eliminated due to the complex stochastic environment in which the system operates.As a result, safety validation is not only concerned about whether a failure can occur, but also discovering which failures are most likely to occur. This article presents adaptive stress testing (AST), a framework for finding the most likely path to a failure event in simulation. We consider a general black box setting for partially observable and continuous-valued systems operating in an environment with stochastic disturbances. We formulate the problem as a Markov decision process and use reinforcement learning to optimize it. The approach is simulation-based and does not require internal knowledge of the system, making it suitable for black-box testing of large systems. We present different formulations depending on whether the state is fully observable or partially observable. In the latter case, we present a modified Monte Carlo tree search algorithm that only requires access to the pseudorandom number generator of the simulator to overcome partial observability. We also present an extension of the framework, called differential adaptive stress testing (DAST), that can find failures that occur in one system but not in another. This type of differential analysis is useful in applications such as regression testing, where we are concerned with finding areas of relative weakness compared to a baseline. We demonstrate the effectiveness of the approach on an aircraft collision avoidance application, where a prototype aircraft collision avoidance system is stress tested to find the most likely scenarios of near mid-air collision.

Verification and Validation↗

Reinforcement Learning in a Nonstationary Environment: The El Farol Problem

This paper examines the performance of simple learning rules in a complex adaptive system based on a coordination problem modeled on the El Farol problem. The key features of the El Farol problem are that it typically involves a medium number of agents and that agents' pay-off functions have a discontinuous response to increased congestion. First we consider a single adaptive agent facing a stationary environment. We demonstrate that the simple learning rules proposed by Roth and Er'ev can be extremely sensitive to small changes in the initial conditions and that events early in a simulation can affect the performance of the rule over a relatively long time horizon. In contrast, a reinforcement learning rule based on standard practice in the computer science literature converges rapidly and robustly. The situation is reversed when multiple adaptive agents interact: the RE algorithms often converge rapidly to a stable average aggregate attendance despite the slow and erratic behavior of individual learners, while the CS based learners frequently over-attend in the early and intermediate terms. The symmetric mixed strategy equilibria is unstable: all three learning rules ultimately tend towards pure strategies or stabilize in the medium term at non-equilibrium probabilities of attendance. The brittleness of the algorithms in different contexts emphasize the importance of thorough and thoughtful examination of simulation-based results.

Bell, Ann Maria↗

Robust integrated neurocontroller for complex dynamic systems

The goal of this research effort is to develop an integrated control software environment for the purpose of creating an intelligent neurocontrol system. The system will be capable of estimating states, identifying parameters, diagnosing conditions, planning control strategies, and producing intelligent control actions. The distinct features of such control system are adaptability and on-line learning capability. The proposed system will be flexible to allow structure adaptability to account for changes in the dynamic system such as sensory failures and/or component degradations. The developed system should learn system uncertainties and changes, as they occur, while maintaining minimal control level on the dynamic system. The research activities set to achieve the research objective are summarized by the following general items: (1) Development of a system identifier or diagnostic system; (2) Development of a robust neurocontroller system, and; (3) Integration of above systems to create a robust Integration Control system (RIC-system). Two contrary approaches are investigated in this research: classical (traditional) design approach, and the simultaneous design approach. However, in both approaches neural network is the base for the development of different functions of the system. The two resulting designs will be tested and simulation results will be compared for better possible implementation.

Zein-Sabbato, S.↗

Robust Integrated Neurocontroller for Complex Dynamic Systems

The goal of this research effort is to develop an integrated control software environment for the purpose of creating an intelligent neurocontrol system. The system will be capable of estimating states, identifying parameters, diagnosing conditions, planning control strategies, and producing intelligent control actions. The distinct features of such control system are: adaptability and on-line learning capability. The proposed system will be flexible to allow structure adaptability to account for changes in the dynamic system such as: sensory failures and/or component degradations. The developed system should learn system uncertainties and changes, as they occur, while maintaining minimal control level on the dynamic system. The research activities set to achieve the research objective are summarized by the following general items: (1) Development of a system identifier or diagnostic system, (2) Development of a robust neurocontroller system, and 3. Integration of above systems to create a Robust Integrated Control system (RIC-system). Two contrary approaches are investigated in this research: classical (traditional) design approach, and the simultaneous design approach. However, in both approaches neural network is the base for the development of different functions of the system. The two resulting designs will be tested and simulation results will be compared for better possible implementation.

Zein-Sabatto, S.↗

Overview of RS-25 Adaptation Hot-Fire Test Series for SLS, Status and Lessons Learned

This paper discusses the engine system design, hot-fire test history and analyses for the RS-25 Adaptation Engine test series, a major hot-fire test series supporting the Space Launch System (SLS) program. The RS-25 is an evolution of the Space Shuttle Main Engine (SSME). Since the SLS mission profile and engine operating conditions differ from that experienced by the SSME, a test program was needed to verify that SLS-unique requirements could be met by the adapted legacy engines. A series of 18 tests, including one engine acceptance test, was conducted from January 2015 to October 2017, to directly support Exploration Mission-1 (EM-1), the first flight of SLS. These tests were the first hot-firings of legacy SSME hardware since 2009. Major findings are described along with top level overview of the engine system.

Vetcha, Naveen↗

Neuromorphic learning of continuous-valued mappings from noise-corrupted data. Application to real-time adaptive control

The ability of feed-forward neural network architectures to learn continuous valued mappings in the presence of noise was demonstrated in relation to parameter identification and real-time adaptive control applications. An error function was introduced to help optimize parameter values such as number of training iterations, observation time, sampling rate, and scaling of the control signal. The learning performance depended essentially on the degree of embodiment of the control law in the training data set and on the degree of uniformity of the probability distribution function of the data that are presented to the net during sequence. When a control law was corrupted by noise, the fluctuations of the training data biased the probability distribution function of the training data sequence. Only if the noise contamination is minimized and the degree of embodiment of the control law is maximized, can a neural net develop a good representation of the mapping and be used as a neurocontroller. A multilayer net was trained with back-error-propagation to control a cart-pole system for linear and nonlinear control laws in the presence of data processing noise and measurement noise. The neurocontroller exhibited noise-filtering properties and was found to operate more smoothly than the teacher in the presence of measurement noise.

Troudet, Terry↗