Search NASA⌕ Search

SEARCH · Search NASA

Results for “Machine Learning (ML)”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 217 records · Page 12

Benchmark Models for Classification of Radiation Type Induced in Immune Cells

NASA Biological and Physical Sciences and the Science Mission Directorate have published a benchmark dataset of mouse immune cells subjected to radiation-induced DNA damage. The dataset comprises ML-ready microscopic imagery of said cells, including labels indicating radiation type and dose. The machine learning team at NASA Interagency Implementation and Advanced Concept Team (IMPACT) created multiple benchmark models. Initially, we conducted a preliminary analysis using thresholding. The algorithm used thresholds on average brightness of the available images to classify them into their respective radiation type. We also tested machine learning approaches. Convolutional Neural Networks (CNN) emerged as the best-performing model. This poster presents the benchmark scores obtained by the models.

Vishal Perekadan↗

Automatic Detection and Classification of Aurora in THEMIS All‐Sky Images

We report a novel machine-learning algorithm for automatically detecting and classifying aurora in all–sky images (ASI) that is largely trained without requiring ground–truth labels. By including a small number of labeled images, we are able to automatically label all of the approximately 700 million images in the Time History of Events and Macroscale Interactions during Substorms (THEMIS) ASI data set from 2008 to 2022. We use a two–stage approach. In the first stage, we adapt the Simple framework for Contrastive Learning of Representations (SimCLR) algorithm to learn latent representations of THEMIS all–sky images. We then finetune a classifier network on the latent representations our model learns of the manually labeled Oslo aurora THEMIS (OATH) data set. We demonstrate that this two–stage approach achieves excellent classification results on data for which there is no current ML classification benchmark. The outcome of this work will facilitate efficient information retrieval for researchers interested in specific categories of aurora and will enable large scale statistical studies and machine learning analyses of THEMIS all–sky images that have not previously been possible. To demonstrate possible ways to utilize this database, we performed a statistical analysis of the occurrence rates of auroral labels with respect to solar wind parameters, interplanetary magnetic field vector, and geomagnetic indices. We further investigate the occurrence rates of auroral phenomena in the annotated data set and their geoeffectiveness by utilizing the co–located THEMIS ground magnetometer data set.

Jeremiah W Johnson↗

Augmenting RANS Turbulence Models Guided by Field Inversion and Machine Learning

This report investigates the use of a data-driven approach, viz., Field Inversion and Machine Learning (FIML), to improve conventional RANS turbulence models like the Spalart-Allmaras model and the Menter SST k-ω model. One of the crucial aspects of using an ML-based approach with limited training data to produce corrections that are generalizable to a large range of flow configurations is to design appropriate “features” (inputs to the ML model). A model, based on guidance from the FIML methodology, is presented in analytical form. An additional list of potential features is provided. Although these were not used in the present correction, they were considered in the course of its development, and are included to fully document the complete process employed in the present work.

turbulence modeling↗

Hazard Contribution Modes of Machine Learning Components

Amongst the essential steps to be taken towards developing and deploying safe systems with embedded learning-enabled components (LECs) i.e., software components that use ma- chine learning (ML)—are to analyze and understand the con- tribution of the constituent LECs to safety, and to assure that those contributions have been appropriately managed. This paper addresses both steps by, first, introducing the notion of hazard contribution modes (HCMs) a categorization of the ways in which the ML elements of LECs can contribute to hazardous system states; and, second, describing how argumentation patterns can capture the reasoning that can be used to assure HCM mitigation. Our framework is generic in the sense that the categories of HCMs developed i) can admit different learning schemes, i.e., supervised, unsupervised, and reinforcement learning, and ii) are not dependent on the type of system in which the LECs are embedded, i.e., both cyber and cyber-physical systems. One of the goals of this work is to serve a starting point for systematizing L analysis towards eventually automating it in a tool.

Smith, Colin↗

Machine Learning for Extravehicular Mobility Unit (EMU) Glove Inspections

The Extravehicular Mobility Unit (EMU) Glove Machine Learning Inspection project utilizes machine learning to expedite the inspection, analysis, and recommendation for continued use of space suit gloves post spacewalks. Today, ISS glove photos are individually reviewed by a team of experts to determine the conditions of space suit gloves. For this project the Microsoft Azure platform is used to perform Automated Machine Learning (AutoML) to detect issues with tagged images from previous Extravehicular Activities (EVA’s) to build a predictive model. The model analyzes a test image and deems the glove GO or NO-GO for additional EVA’s. The goal for this ML project is to decrease the time spent reviewing images by ground personnel and crewmembers in high frequency EVA locations such as the Moon and Mars. For destinations such as the Moon and Mars the goal is to give crew autonomy in determining glove conditions with limited support from Earth. This paper will outline the results to date and future work needed to expand the capability for in-situ recommendations.

EVA↗

Prediction of Aircraft Estimated Time of Arrival Using A Supervised Learning Approach

We present a novel data-driven approach for prediction of the estimated time of arrival (ETA) of aircraft in the terminal area via the implementation of a Random Forest regression model. The model uses data fused from a number of sources (flight track, weather, flight plan information, etc.) and provides predictions for the remaining flight time for aircraft landing at Dallas/Fort Worth (DFW) International Airport. The predictions are made when the aircraft is at a distance of 200-miles from the airport. The results show that the model is able to predict estimated time of arrival to within ± 5 min for 90% of the flights in the test data with the mean absolute error being lower at 145 seconds. This paper covers the entire pipeline of data collection, preprocessing, setup and training of the ML model, and the results obtained for DFW.

Machine learning↗

Data-Driven Study of Shape Memory Behavior of Multi-component Ni-Ti Alloys

Ni-Ti based shape memory alloys (SMAs) have found wide-spread use in aerospace, automotive, biomedical, and commercial applications owing to their favorable properties and ease of operation. Especially important for many NASA applications is the ability to tune the martensitic transformation temperature of Ni-Ti alloys by varying the composition and processing conditions. Recently, researchers at NASA have compiled an extensive database of shape memory properties of materials, including over 8,000 multi-component Ni-Ti alloys containing 37 different alloying elements. Using this dataset, machine learning models are trained to predict transformation temperatures, hysteresis, and transformation strain with extremely small errors. These models are used to learn relationships between shape memory behavior and input parameters in the composition and processing space. ML predictions are validated through new experiments. The combination of an extensive dataset and accurate learning models, together, make our approach highly suitable for the rapid discovery of novel SMAs with targeted properties.

Shape Memory Alloys↗

Data-Driven Study of Shape Memory Behavior of Multi-Component Ni-Ti Alloys

Ni-Ti based shape memory alloys (SMAs) have found wide-spread use in aerospace, automotive, biomedical, and commercial applications owing to their favorable properties and ease of operation. Especially important for many NASA applications is the ability to tune the martensitic transformation temperature of Ni-Ti alloys by varying the alloy composition and processing conditions. Recently, researchers at NASA have compiled an extensive database of shape memory properties of materials, including over 8,000 multi-component Ni-Ti alloys containing 37 different alloying elements. Using this dataset, machine learning models are trained to predict transformation temperatures, hysteresis, and transformation strain with extremely low mean absolute errors. These models are used to learn relationships between shape memory behavior and input parameters in the composition and processing space. ML predictions are validated through new experiments. The combination of an extensive experimental dataset and accurate learning models, together, make our approach highly suitable for the rapid discovery and design of novel SMAs with targeted properties. We are not aware of any current approaches capable of predicting SMA transformation behavior over such a wide range of compositions and processing conditions.

Shape memory alloys↗

Parameterization of Vertical Cloud Distribution from C3M and MERRA Data Using ML Method

Clouds play a key role in regulating the hydrological cycle and the Earth's radiative energy budget. However, global climate models (GCMs) with a horizontal grid spacing on the order of 100 km have limitations in representing sub-grid cloud dynamics with spatial scales on the order of 1 km, leading to potential uncertainties in cloud radiative feedback on the global scale. In our research, we will leverage the capabilities of Deep Machine Learning (DML) methods to construct parameterizations of sub-grid volumetric cloud fraction (VCF), which is the frequency of occurrence on a grid volume accumulated in the horizontal and vertical directions. Our investigation delves into the intricate relationship between VCF obtained from the NASA CALIPSO-CloudSat-CERES-MODIS (CCCM) satellite observation data and 3-D MERRA-2 reanalysis meteorological profiling data (e.g., wind, relative humidity, temperature). Through a comprehensive one-year data training utilizing the Sequence to Sequence DML method, we have successfully disentangled the complicated cloud formation dynamics across diverse meteorological conditions through a day-to-day analysis framework. Preliminary findings reveal promising statistical agreements in geographical and vertical distributions and seasonal variations of volumetric cloud fraction between ML prediction and satellite measurements. These results underscore the aptitude of our DML model to discern underlying cloud physical processes and accurately represent sub-grid cloud formation dynamics. Additionally, we have also employed trained neural network to analyze uncertainties arising from errors in meteorological data, further enhancing the robustness of our VCF parameterization.

Shan Zeng↗

Nearest-Neighbor Machine Learning Feature Selection for Interpretation of Microbial Molecular Signatures from Isotope Ratio Mass Spectrometry Data

Mass spectrometry (MS) promises to be a powerful tool for potential biosignature detection during astrobiological missions on ocean worlds in our solar system. Accurate and generalizable machine learning methods could enhance science return on investment by predicting seawater chemistry and classifying isotopic biosignatures, either as a signature consistent with microbial life (biotic) or as a novelty (unclassified/unique). However, machine learning models are likely to be complex and involve interactions between MS features, making biosignatures difficult to interpret. Feature selection methods provide biological and chemical context that help interpret the mechanisms of machine learning models, but these methods also need the ability to detect complex interactions. Previously, we developed a machine learning feature selection algorithm called nearest-neighbor projected distance regression (NPDR) that has the ability to identify important model features that involve complex interactions and automatically reduce correlation and the dimensionality in a high-dimensional variable space. The standard distance metrics used in NPDR – Manhattan and Euclidean – assume the multivariate data are isotropic, which is often violated in real data due to differences in the covariance between variables. Thus, we extend NPDR to include a random forest distance, and other anisotropic distance metrics, for computing nearest neighbors. We also augment the isotope-ratio MS data with time-series features from the raw MS signal to improve biotic classification. We test NPDR on our novel experimental ocean world seawater analog MS data. We measure isotope fractionations of volatile CO 2 that could be measured in exospheres or plumes. Samples include baseline abiotic conditions using a range of possible seawater chemistry consistent with Europa and Enceladus, and biotic samples that include microbes in these seawaters. We use penalized NPDR with random forest proximity to identify interpretable microbial molecular signatures. We compare features with random forest importance, and we train a classifier that discriminates between biotic and abiotic samples with high accuracy. These ML-trained ocean-world analog MS data could be used to assist in identifying biosignatures during future missions.

geochemistry↗

Report on Workshop on Artificial Intelligence in Strategic Planning and Science Prioritization

This report details the observations from a two-day virtual workshop, held May 12-13, 2020, focused on whether, and how, artificial intelligence (AI) could assist humans in strategic planning, specifically in science and technology prioritization. The participants identified several “key challenges” that AI might tackle in this area. To further understand the value of these key challenges the workshop then developed related test cases that would demonstrate specifically how AI/machine learning (ML) could provide assistance to humans. Approximately 40 subject matter experts (SMEs), with backgrounds in AI, strategic planning for science, and scientific data, were gathered for the conference. This report collates the details of the output of the workshop. The “best” test cases include (in no particular order):Use of AI to assist in selecting Decadal Survey priorities. * Use of AI to identify new, or previously unidentified, science topics for prioritization. * Using AI to better label and increase discoverability of scientific literature and proposals. * Use of AI to enhance current observation capabilities for scientific missions. * Using AI to mitigate biases in selection of proposal reviewers and membership of advisory committees. Examination of these test cases indicates that Natural Language Processing (NLP) is a common capability found in most of the ”best” (top-rated) test cases and is a valuable, multi-purpose tool which enables ML in this area.

strategic planning↗

Coaxial Rotor CFD Validation and ML Surrogate Model Generation

A fixed-pitch speed-controlled coaxial rotor system was tested in the NASA Langley Transonic Dynamics Tunnel in September 2022. The rotors have a diameter of 1.35 meters and an inter-rotor spacing of 25% of the diameter. Though this test was focused on the NASA New Frontiers Dragonfly mission, the resulting dataset is relevant to a wide array of applications of multirotor vehicles, especially those using fixed-pitch variable-speed rotors. Most notably beyond Dragonfly, perhaps, is the application of coaxial rotor pair configurations for eVTOL and Urban Air Mobility (UAM) aircraft. This work provides a thorough CFD validation study quantifying coaxial rotor performance estimation with accuracy on order 5-10% using an efficient hybrid BEMT-URANS flow solver over a wide range of operating conditions. This accuracy was achieved using novel approaches for the construction of both the C81 airfoil performance lookup tables and the BEMT rotor model. These novel approaches are combined with advanced scripting to further accelerate the commercial off-the-shelf CFD solver on GPU accelerated machines. Finally, the development of machine learning surrogate models is presented to produce highly efficient and accurate rotor performance predictions for fixed-pitch variable-speed multirotor aircraft over the complete flight regime including scout, cruise, climb, descent, as well as limiting cases of vortex-ring and windmill-brake states.

Coaxial↗

Recommendations on Evidence and Process for Certification of Learning-enabled Components in Aerospace Systems

This report primarily identifies a collection of relevant and necessary evidence for assurance of machine learnt components (MLCs)—also known as learning-enabled components—integrated into aircraft systems, and gives preliminary suggestions on the elements of a certification process that invoke the identified evidence. The main focus is on feedforward neural networks that are static and trained offline through supervised learning. A brief background on the generic elements of the lifecycle of an MLC is given to contextualize the assurance considerations and, consequently, the evidence that is relevant and necessary to support certification. At the level of an MLC, those considerations relate to: (i) the consistency and correctness of MLC contributions to system functions in the context of a validated functional intent; and (ii) the absence of MLC contributions to aircraft-level failure conditions. At an ML model level, confidence in model and data properties contribute to assurance of the containing MLC, in particular: (a) generalizability and robustness of models, in the presence of inputs not previously seen during training, disturbances to inputs, and unexpected inputs; and (b) valid data, i.e., data that are at least representative, relevant, complete, and accurate. Evidence for the above span the elements of the ML lifecycle, and includes, at a minimum, lifecycle artifacts that pertain to: (1) properties of requirements capturing functional intent, safety constraints, and aspects of the intended use and operating environment; (2) model performance, model complexity and design, and algorithm choice; (3) achievement of required performance at the levels of a trained model during model development, a trained model after model development is complete, and a trained model that is transformed into an executable equivalent; (4) model implementation aspects necessary for transforming a trained model into the executable equivalent; (5) integration of the executable trained model into the containing MLC, and eventually the larger system; and, (6) lastly, the verification and validation (V&V) of each of the above. Such V&V lifecycle artifacts themselves include: aspects of coverage, e.g., of various levels of requirements by the input space of the model and the data; traceability (where applicable); application of formal methods for property specification, analysis, and checking. Examples of evidence generation methods and tools further ground the discussion on what constitutes evidence, and the contribution to assurance during certification. The identified assurance considerations and supporting evidence is not a comprehensive set. Additionally, neither what should be considered as sufficient evidence relative to the assigned criticality of an MLC, nor how criticality ought to be determined and adjusted, have been considered in this report. However, suggestions are made for potential activities of the ML lifecycle that are aimed at providing confidence that an MLC can be relied upon when integrated into its containing (aircraft) system. Those activities are proposed as candidate elements of a certification process for MLCs. The main purpose of this report to inform regulatory guidance and consensus standards that may be used to meet the safety intent of the applicable regulations.

Aviation safety↗

A New ML-Based Adaptive Thinning Methodology to Improve the Impact of AIRS and CrIS Assimilation on Global Tropical Cyclone Forecasts

This work builds on previous research performed by this team to improve the forecast of Tropical Cyclones (TCs) by assimilating AIRS and CrIS radiances into the NASA Global Earth Observing System (GEOS). Past published work demonstrated that the assimilation of radiances with variable density was beneficial to TC forecasting in the GEOS. In the previous setup, a fixed-size moving square named 'TC domain' was activated by the so-called TC-vitals, an international real-time message accessible to all NWP forecasting centers, that documents the existence of a TC, its estimated position, and its size. The information from TC-vitals activated a switch in the GEOS, which allowed to reduce the distance used for thinning AIRS and CrIS data inside a 15 degrees by 15 degrees moving TC domain centered on the storm, so that more data were assimilated in the vicinity of the TC during its lifetime. The methodology produced improved TC analyses and led to better forecasts, particularly related to intensity, without damaging the global forecast skill. In the new version, the adaptive thinning methodology is based on a machine-learning technique. The technique searches for TCs and creates TC masks by using cloud-top temperatures from all geostationary satellites without the need for additional information. It is being trained against the International Best Track Archive for Climate Stewardship (IBTrACS) data base. Once a TC mask is created, a switch identical to the one used in the previous adaptive thinning method is activated, allowing the GEOS to ingest more data in the TC-shaped size-changing domain that follows the storm. As of today, the team has been able to successfully assimilate data inside the ML-detected TC domains. Future work includes an improved capability of reducing false alarm rates (i.e., cloud systems that are erroneously labeled as TCs).

Oreste Reale↗

Interpretable ML Approaches for Novel Solid State Electrolyte Design

All-solid-state batteries with Li metal anode can address the safety issues surrounding traditional Li-ion batteries as well as the demand for higher energy densities. However, the development of solid electrolytes simultaneously possessing high ionic conductivity and good chemical and electrochemical stabilities has proven to be a challenge. I will present our informatics approach to explore the Li compound space for promising solid electrolytes using high-throughput multi-property screening and interpretable machine learning. This is accomplished through the generation of a large database of battery-related materials properties of Li compounds. We use tree-based ensemble learning methods and graph neural network approaches to accurately learn relationships between crystal structures and corresponding thermodynamic and kinetic properties, with interpretability being a major focus. Our models give us the ability to enable rapid discovery and design of novel solid-state battery chemistries.

Shreyas J Honrao↗

Ground Stop Adjuster: A Machine Learning Approach to Improve Air Traffic Management Initiatives

Traffic Management Initiatives (TMIs) play a crucial role in balancing demand and capacity within the U.S. National Airspace System (NAS). In current practice, traffic management coordinators (TMCs) determine and issue TMIs and recent research has explored the use of machine learning tools to aid the TMCs. However, most studies have primarily focused on a particular type of TMI, i.e., Ground Delay Programs (GDPs) due to their higher rate of occurrence and longer duration. This study investigates a machine learning approach for monitoring and adjusting a different type of TMI, i.e., Ground Stop (GS), aiming to assist human decision-makers with accurate, consistent, and timely recommendations. Using data from three major airports in the New York metroplex, we evaluated models that predict GS parameters, such as duration and scope. Our results demonstrate that using data from all airports in the NY metroplex and increasing feature granularity improve the prediction accuracy of the ML models.

Farzan Masrour Shalmani↗

A Machine Learning Ready Dataset of Acoustic Power Maps for Detection of Active Region Emergence

The development of an accurate forecast for solar eruptive activity has become increasingly important in order to prevent any potential impact on activities in space and the Earth's environment. It is therefore crucial to detect active regions before they appear on the solar surface and create early warning capabilities for upcoming Space Weather disturbances. In this work, 9TB of solar data (SDO/HMI dopplergrams, magnetograms and continuum intensity maps) involving the emergence of 61 NOAA solar active regions since 2010 were processed using the NASA HECC capabilities. An acoustic power maps time-series dataset was created (for four different frequency ranges and processed to take into account the solar sphere geometric effect ) which can be used for understanding the dynamics of the solar surface and train a variety of ML models. The calculated acoustic power maps carry precursor information associated with the decrease in continuum intensity on the solar surface, verifying older helioseismology research. Our results show that a Long Short-Term Memory (LSTMs) model, with a modest layer depth and the right hyperparameters tuned, when trained on this solar acoustic power maps dataset can predict without false negatives a drop in intensity (associated with the emergence of the active region), up to 18 hours in advance.

SMD↗