Search NASA⌕ Search

SEARCH · Search NASA

Results for “Learning and learning models”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 181 records · Page 10

ES2Vec: Earth Science Metadata Suggestions and Analogical Reasoning

As the volume of text-based Earth science research grows, it is increasingly possible to discover latent relationships in the literature. However, traditional methodologies are restricted by limited computational capabilities and intractable problem spaces. Advancements in natural language processing (NLP) have allowed us to use an extensive Earth science corpus to create a domain-specific word vector model, Es2Vec, which we have used to surface latent relationships between Earth science concepts and generate improved keyword tags. Earth science metadata keyword assignment is a challenging problem. Dataset curators select appropriate keywords from the Global Change Master Directory (GCMD) set of keywords. The keywords an are integral part of the search and discovery of these datasets. Hence, the selection of keywords is crucial to increasing the discoverability of datasets. Utilizing machine learning techniques, we provide users with automated keyword suggestions to complement manual selection. We trained a machine learning model that leverages the semantic embedding ability of Word2Vec models to process abstracts and suggest relevant keywords. A user interface tool we built to assist data curators in the assignment of such keywords is also described.

word vectors↗

An Artificial Neural Network Approach to Predict Rotor-Airframe Acoustic Waveforms

A surrogate artificial neural network/machine learning model was developed to predict the acoustic interaction for a fixed-pitch rotor in proximity to a downstream cylindrical airframe typical of small Unmanned Aerial System (sUAS) platforms. The model was trained to predict the acoustic waveform under representative hover conditions as a function of rotational speed, airframe proximity, and observer angle. Training data were acquired in an anechoic chamber on both isolated rotors and rotor-airframe configurations. Acoustic amplitude and phase of the revolution-averaged interaction were predicted, which required up to 25 harmonics to capture the impulse event caused by the blade’s approach and departure from the airframe. Prediction performance showed, on average, that the models could estimate the acoustic amplitude and phase over the relevant harmonics for unseen conditions with 86% and 75% accuracy, respectively, enabling a time domain reconstruction of the waveform for the range of geometric and flow parameters tested.

acoustics↗

An Artificial Neural Network Approach to Predict Rotor-Airframe Acoustic Waveforms

A surrogate artificial neural network/machine learning model was developed to predict the acoustic interaction for a fixed-pitch rotor in proximity to a downstream cylindrical airframe typical of small Unmanned Aerial System (sUAS) platforms. The model was trained to predict the acoustic waveform under representative hover conditions as a function of rotational speed, airframe proximity, and observer angle. Training data were acquired in an anechoic chamber on both isolated rotors and rotor-airframe configurations. Acoustic amplitude and phase of the revolution-averaged interaction were predicted, which required up to 25 harmonics to capture the impulse event caused by the blade’s approach and departure from the airframe. Prediction performance showed, on average, that the models could estimate the acoustic amplitude and phase over the relevant harmonics for unseen conditions with 86% and 75% accuracy, respectively, enabling a time domain reconstruction of the waveform for the range of geometric and flow parameters tested.

acoustics↗

Natural Language Understanding and Extraction of Flight Constraints Recorded in Letters of Agreement

This paper presents an automated information extraction and inference technique using natural language processing for extracting flight operational procedures and constraints embedded in heritage air traffic management documents. The extracted flight constraints can be digitized and fit into existing airspace information exchange models such as the Aeronautical Information Exchange Model (AIXM). This approach offers a digitized solution to disseminate airspace operating conditions to diverse air users and stakeholders in the National Airspace System (NAS). Furthermore, the digitized flight procedures can provide operational flexibility for emerging advanced air mobility providers and reduce traffic controller workload while maintaining current safety standards. To demonstrate this process, 1,972 Letters of Agreement (LOAs) have been selected for processing, named entity extraction, constraint identification and extraction. This dataset is derived from a subset of documents related to Air Route Traffic Control Centers (ARTCC) operations. We experimented with various traditional information extraction techniques, state-of-the-art machine learning and deep learning models to perform named entity recognition and pattern recognition on our dataset. We present the results from our experiments and demonstrate 99.0% F-1 score for named entity recognition, and a 96.6% accuracy for our entire workflow up to named entity recognition. We also discuss constraint definitions using generic patterned templates and extensions to this work in applying entity linking to digitally extracting relevant constraints.

Natural Language Processing↗

Natural Language Understanding and Extraction of Flight Constraints Recorded in Letters of Agreement

This paper presents an automated information extraction and inference technique using natural language processing for extracting flight operational procedures and constraints embedded in heritage air traffic management documents. The extracted flight constraints can be digitized and fit into existing airspace information exchange models such as the Aeronautical Information Exchange Model (AIXM). This approach offers a digitized solution to disseminate airspace operating conditions to diverse air users and stakeholders in the National Airspace System (NAS). Furthermore, the digitized flight procedures can provide operational flexibility for emerging advanced air mobility providers and reduce traffic controller workload while maintaining current safety standards. To demonstrate this process, 1,972 Letters of Agreement (LOAs) have been selected for processing, named entity extraction, constraint identification and extraction. This dataset is derived from a subset of documents related to Air Route Traffic Control Centers (ARTCC) operations. We experimented with various traditional information extraction techniques, state-of-the-art machine learning and deep learning models to perform named entity recognition and pattern recognition on our dataset. We present the results from our experiments and demonstrate 99.0% F-1 score for named entity recognition, and a 96.6% accuracy for our entire workflow up to named entity recognition. We also discuss constraint definitions using generic patterned templates and extensions to this work in applying entity linking to digitally extracting relevant constraints.

Natural Language Processing↗

Presound: UAV Diagnostic System Enabled by Vibration-Based Machine Learning

A low-weight, inexpensive small unmanned aerial system (sUAS) that takes off, performs a mission, lands, and safely stows and recharges itself has myriad future applications ranging from agricultural imaging to last-mile package delivery. Likewise, Urban Air Mobility (UAM) systems will enable people to take air taxis from point to point in cities, rapidly moving commuters long distances without concern for road traffic and congestion. Fully electric aviation systems will be cleaner and quieter than ground transport. Cities could eliminate cars and buses, and convert roads to higher capacity bike and pedestrian throughways. Yet, for sUAS as well as UAM, system reliability and assurance is a limiting factor to deploying affordable autonomous flight systems. For this bright future of aviation to be realized, aircraft must be able to autonomously and accurately self-diagnose health issues both before takeoff and during flight. The GreenSight PreSound system is designed to identify defects on aircraft through intelligent analysis of vibration. It accomplishes this by measuring structural vibrations induced by the vehicle’s own propellers, and analyzing that data using a machine learning model that determines whether a defect is present. The PreSound system is designed to require no human oversight, and to operate across a wide array of vehicles through re-training of the model for each target aircraft. PreSound has been developed and seen limited early success using data collected from the GreenSight Dreamer sUAS, a 5lb quadrotor vehicle designed for aerial imaging applications. The final detection model, trained on data with props spinning at 50% throttle, achieves excellent performance with over 99% average accuracy in detecting blade damage using a single FFT vector input. It demonstrates the ability to generalize to new types of blade damage, correctly classifying a different type of blade damage with 98% accuracy. Full test pulses were classified with 100% accuracy, and in live testing, all sets of data during blade movement were classified accurately with over 95% confidence. When trained on in-flight data, the same model achieves an average accuracy of 85% in distinguishing between undamaged and blade-damaged states in flight. The authors believe that these accuracies show significant potential of this approach to expand unmanned flight safety, with significant potential benefits in accelerating Advanced Aerial Mobility (AAM) and UAM aviation applications.

UAS↗

Utilizing Earth Observations to Understand Landscape Patterns and Assist in Wildlife Management in Iona National Park, Angola

Following the end of the Angolan Civil War (1975-2002), human habitation in Iona National Park has grown exponentially, as has the livestock population. An ongoing drought beginning in 2017 has brought people, livestock, and wildlife into increasing competition for resources within the park. This study used Earth observation data, primarily Landsat and Sentinel imagery, to examine landscape trends to improve wildlife preservation approaches in Iona National Park, Angola. In collaboration with the NGO African Parks, we developed a robust land use and land cover (LULC) classification model using remote sensing data to augment sparse ground-based data in this arid land region. We used Google Earth Engine and a random forest classifier to map vegetation types, water bodies, and potential wildlife habitats. This analysis resulted in a high spatial resolution LULC time-series between 1984-2023, highlighting critical periods of socioecological change over the past 40 years. These results increased the partner’s ability to make scientifically grounded decisions about resource allocation and conservation priorities. This analysis supports the feasibility of applying remote sensing techniques coupled with machine learning models in dry regions, where standard survey methods are frequently limited by accessibility and resource availability. However, we identified limitations in ground-truth data and the difficulty of recognizing certain vegetation types in arid areas. Despite these limitations, the study demonstrated Earth observations' ability to transform wildlife management techniques in distant and data-scarce locations, providing a reproducible foundation for similar ecosystems around the world.

Emmanuel Aklie↗

Advancements in Blowing Dust Detection at Night via Machine Learning

This presentation introduces operational users to a machine-learning based Dust Probability product developed by the NASA SPoRT program for the application of detecting and monitoring blowing dust plumes at night. Advances in earth observing satellites has improved monitoring and detection of dust both day and night through derived imagery such as the Dust RGB. However, limitations of the RGB at night result in less contrast between dust and land surface features, as seen by the user. A Machine Learning (ML) model has been developed and applied to GOES-16 ABI to overcome this limitation and improve nighttime dust detection. The ML capability is a subset of Artificial Intelligence methods. In this case the Dust ML model was developed using a simple Random Forest (RF) model, typically used to solve classification challenges (or to provide regression type output). The goal was to leverage the strengths of the RF model to learn how to identify blowing dust, and hence, overcome the limitation of a user trying to detect blowing dust within the satellite imagery by eye alone. A brief description of the ML model development will be provided. However, the focus of the presentation will be on the initial user feedback from the assessment of this tool for the 2022 blowing dust events of March through April. During this time several users across the U.S. Southwest collaborated to apply this Dust ML product at night as a complement to the existing Dust RGB in order to determine if it provided greater operational efficiency and value.

Machine Learning↗

Developing Open-Source Training Materials for AI/ML and Space Biological Sciences Using NASA Cloud-Based Data

Artificial Intelligence (AI) and Machine Learning (ML) has gained significant traction in the biological and biomedical research fields, in part due to a culture of open data sharing and reuse. AI/ML methodology is well-suited to recognize and predict biological patterns from high-dimensional next-generation sequencing data (e.g. whole genome sequencing, transcriptomic sequencing), as well as from biological or medical imaging data (e.g. microscopy, computed tomography, ultrasound, magnetic resonance imaging, radiography). These methodologies hold particular promise for space biosciences research and automated space health monitoring systems. However, there are key considerations for properly training, validating, and testing a machine learning model in biological research or clinical application. Inexperienced researchers can produce models that perform poorly outside of the training dataset. Open Science principles such as data sharing and open-source code must go hand-in-hand with publicly available, high-quality training curricula in best practices, with modules centered on real-life scientific use cases and data so future AI/ML practitioners gain experience on real problems. Here we present the development of open-source training materials for AI/ML and space biosciences, as part of the NASA Transform to Open Science Training (TOPST) initiative. We develop 4 independent training programs, focused on the following topics: 1) Fundamentals of Machine Learning and Space Biosciences Domain, 2) Open Science, Artificial Intelligence, and Ethical Best Practices for Data Sharing and Analysis, 3) Using AI/ML Classification to Identify Gene Networks Affected By Space Exposure in Mouse Liver, and 4) Using Neural Networks to Find DNA Damage Patterns in Immune Cells after Radiation. All programs leverage cloud-based NASA biological datasets. The curriculum we present will enable worldwide access to training in AI/ML and scientific analysis.

James Casaletto↗

NASA Satellites and Data Fusion: A Case Study with Coral Reefs

As the world is experiencing a significant rise in both AI, Climate, and Space start-up companies, we are in a new wave of limitless innovation. NASA's statutory responsibility is to "provide for the widest practicable and appropriate dissemination of information concerning its activities and the results thereof." (51 U.S.C. § 20112) In particular, through machine learning, the public data drawn from NASA’s space assets can provide insights for addressing climate-related problems here on Earth. And many climate start-up companies can benefit from leveraging this data, either for proof of concepts or their own missions. In the Fall of 2021, several scientists from NASA and Coral Vita led a Practicum with the Georgia Institute of Technology Masters in Data Analytics program. The Practicum saw two teams of students develop and implement machine learning models to infer, from CALIPSO satellite imagery, vitality properties upon satellite pass. This kind of capability can lead to real-time mapping of coral health around the world, giving organization an understanding of where to prioritize reef reconstitution.. In this presentation, we will highlight this use case and discuss other concrete applications in which space data is used to solve problems here on Earth.

Earth Sciences↗

Variance Decomposition of MEDLI2 Reconstructed Heating Using Neural Networks

The Mars Entry, Descent, and Landing Instrumentation (MEDLI2) sensor suite collected data during entry of the Mars 2020 Perseverance rover into Mars’ atmosphere. This suite included a network of MEDLI2 Instrumented Sensor Plugs (MISPs). Each MISP was comprised of a cylinder made of Thermal Protection System (TPS) material with 1-3 embedded thermocouples (TCs), and it was flush mounted into the heatshield or backshell. Data from these in-depth TCs were used to reconstruct the aeroheating environment of the vehicle throughout entry. Surface heating was posed as an inverse problem, with the goal of estimating the surface heating by minimizing an objective function of the difference between MISP temperature measurements during flight and the temperature predictions derived from the Fully Implicit Ablation and Thermal response (FIAT) program. Given an aerothermal environment, FIAT calculates the material response and provides in-depth temperatures throughout the TPS material. To achieve the reverse, an internal tool called FIAT_Opt runs through multiple different environments until the output temperature at the TC depth closely matches the flight data. 95% confidence intervals on the reconstructed surface heating were obtained using Monte Carlo analysis, in which uncertainties in the thermocouple depth and the TPS material properties (e.g., density, thermal conductivity, heat capacity, emissivity) based on flight-lot material testing were included. A variance decomposition method using Sobol indices was employed to assess the sensitivity of the reconstructed peak heating to the TC placement and material property uncertainties. Variance decomposition was found to require tens of thousands of FIAT_Opt runs in order for the Sobol indices to converge. With a single FIAT_Opt run taking on the order of 40 minutes, the required number of computations would take months to complete, even if using multiple CPUs. To mitigate this problem, three machine learning models (ridge regression with cross-validation, random forest regression, and a deep neural network) were trained and tested using the 2000 Monte Carlo runs that were already completed. A subset of 1600 runs were used to train the model (i.e., training set), while the remaining 400 runs were used as the test set. The predictions from the deep neural network (DNN) on the test set showed nearly perfect agreement to the actual values computed with FIAT_Opt (R2 > 0.99). Using the DNN as a surrogate model, the variance decomposition using 50,000 runs was completed within minutes. The resulting Sobol indices showed that the reconstructed peak surface heating was most sensitive to the uncertainties in the thermal conductivity (ST = 0.37) and heat capacity (ST = 0.26). This method can be leveraged to provide requirements for material property measurements needed to improve the accuracy of surface heating prediction and ultimately lead to the reduction of design margins in the future. This presentation will include background on the MEDLI2 suite; the method used for inverse heating estimation; the way that material property uncertainties were accounted for using Monte Carlo analysis; a brief background on variance decomposition; the motivation for using machine learning in this context; how a neural network was trained on the data to enable variance decomposition in a fraction of the time; and the variance decomposition results for one of the MISPs.

Hannah Alpert↗

Do Better Satellite Precipitation Algorithms Improve Landslide Hazard Assessment?

Satellites make it possible to estimate precipitation in near real time. Given the challenges of achieving global coverage by other means, these data are used widely. However, few systems for landslide hazard assessment rely on satellite precipitation estimates. This could be due in part to perceptions of accuracy, although latency, spatial resolution, and other factors may also be important. We test whether recent changes to data streams from the Global Precipitation Measurement mission (GPM) have improved its potential for use in landslide prediction. Specifically, we examine data produced by the Integrated Multi-satellitERetrievals for the GPM (IMERG) algorithm, which was upgraded to version 7 this year. IMERG relies upon other algorithms, including the Goddard Profiling Algorithm (GPROF) and the GPM Combined Radar-Radiometer Algorithm (CORRA). Many changes have been made during the switch from IMERG version 6 to version 7. These include upgrading CORRA and GPROF to version 7, to improve the accuracy of precipitation in frozen, mountainous, and coastal areas. The measured intensity of some storms has been enhanced with a new algorithm, the Scheme for Histogram Adjustment with Ranked Precipitation Estimates in the Neighborhood. Combined with many others, these changes to IMERG should improve its utility for landslide hazard assessment in a variety of contexts. To test this idea, we retrain the global Landslide Hazard Assessment for Situational Awareness (LHASA) model twice—first with data from IMERG version 6B and second with 7B. Since current daily rainfall is the most important variable in determining outcomes predicted by LHASA, it should reflect changes made to that input. First, we grid the landslides at a daily, thirty-arcsecond resolution. This serves as the response variable. At each of these sites current and antecedent rainfall are extracted, along with antecedent snow mass and soil moisture, slope, and PGA. In addition, one million grid cells are selected at random points to represent conditions under which landslides (probably) do not occur. After merging these data, we hold back 20% of the dataset for validation purposes and train a machine-learning model with the rest. We assess both the model’s overall ability to identify landslides and its ability to predict specific large landslide disasters.

Thomas A Stanley↗

Satellite Based Precipitation Estimation in Orographic Regions within the Southwestern United States

Predicting precipitation-induced landslides requires accurate estimation of orographic precipitation. Many research studies have been done to estimate orographic precipitation using measurement/estimation methods that include rain gauges, ground-based radar, satellite-based estimates, and modeling. Each method has strengths and weaknesses, but none have been able to fully resolve orographic precipitation. The Integrated Multi-satellitE Retrievals for Global Precipitation Measurement Mission (IMERG) early run product provides precipitation estimates at a 0.1° spatial resolution, at half-hour time scales, with only a 4-hour latency. This product is currently used for the global landslide hazard assessment, but there are known issues in mountainous terrain that the IMERG algorithm has not been able to fully resolve; and the sparse gauge density in these regions makes it even more difficult. In this study, precipitation events were identified using high temporal resolution (5-15 minute) precipitation observations from rain gauges in mountainous terrain in the southwestern United States. The brightness temperature from several infrared (IR) bands from the Geostationary Operational Environmental Satellite (GOES) 16 satellite were used to estimate precipitation with a k-nearest neighbor machine learning model. Compared to IMERG, the IR estimates from GOES-16 performed better at predicting gauge-identified precipitation events. Additionally, the IR-only based estimates were able to estimate precipitation when IMERG failed to detect precipitation, false negative events. While additional analysis is needed, results indicate the need for better integration of IR observations for more accurate precipitation estimation in mountainous regions.

Jessica Sutton↗

Machine Learning for Dynamical Modeling of a Flexible Inverted Pendulum System

The inverted pendulum system, a canonical example of an unstable mechanical system, is often used to model the control problems encountered in the flight of rockets in the initial stages of launch, when the airspeed is too small for aerodynamic stability. A system with a flexible pendulum is a variant that more accurately simulates rocket flight nonlinearities (particularly, the flex modes of the rocket). To increase NASA capability of modeling dynamical systems for which closed form solutions are not clear or easily developed, this project aims to provide a machine learning approach that produces a learned dynamical model of the PENNY robot (a rover with a flexible inverted pendulum) from operational data. The developed approach can then be generalized to other complex dynamical systems, including but not limited to rockets and other robotic systems.

M. A. DuPuis↗

Deep Learning Vetting of TESS FFI Data: Results and Comparison with 2-Min Data

We present the results of vetting TCEs from the TESS SPOC full-frame images (FFI) Year 5 data using our deep learning model, and we compare the performance in this dataset against the results obtained for the TESS SPOC 2-min data. The 200-second cadence FFI data expands the search to a list of targets that not only includes 2-minute targets, but also potentially high-value targets within 100 parsecs or with H-magnitude <10, and field targets with TESS magnitude <13.5. This work aims to explore this rich dataset and increase the efficiency and throughput of the vetting process by helping unearth more high-quality planet candidates from the TESS mission.

tess spoc↗

Validation of Machine Learning Algorithms for Hyperspectral Inversion of Common Water Quality Indicators

The upcoming transition to a diverse suite hyperspectral airborne and orbiting optical sensors will provide an unprecedented opportunity to measure inland water quality characteristics at a fidelity not previously achievable. This presentation will assess prototype deep learning models trained on synthetic hyperspectral data and validated with collocated in-situ measurements. Synthesized data is becoming increasingly popular for use in data-driven approaches to complex problems, and can compliment real data to increase performance on complex and unusual phenomenon, reduce or test bias, and experiment to demonstrate explainability. We will present insights from hyperspectral inversions of Chlorophyl-a, Phycocyanin, and concentration of non-algal particles using selected orbiting and airborne sensors over diverse, optically complex aquatic scenarios. We analyze how various optical water types affect fidelity of results and where improvements can be made as we prototype for globally operational water quality algorithms which can be leveraged by upcoming hyperspectral missions such as the Surface Biology and Geology (SBG) mission.

Surface Biology and Geology (SBG)↗

Low-Cost Sensor Performance Intercomparison, Correction Factor Development, and 2+ Years of Ambient PM2.5 Monitoring in Accra, Ghana

Particulate matter air pollution is a leading cause of global mortality, particularly in Asia and Africa. Addressing the high and wide-ranging air pollution levels requires ambient monitoring, but many low- and middle-income countries (LMICs) remain scarcely monitored. To address these data gaps, recent studies have utilized low-cost sensors. These sensors have varied performance, and little literature exists about sensor intercomparison in Africa. By colocating 2 QuantAQ Modulair-PM, 2 PurpleAir PA-II SD, and 16 Clarity Node-S Generation II monitors with a reference-grade Teledyne monitor in Accra, Ghana, we present the first intercomparisons of different brands of low-cost sensors in Africa, demonstrating that each type of low-cost sensor PM2.5 is strongly correlated with reference PM2.5, but biased high for ambient mixture of sources found in Accra. When compared to a reference monitor, the QuantAQ Modulair-PM has the lowest mean absolute error at 3.04 μg/m3, followed by PurpleAir PA-II (4.54 μg/m3) and Clarity Node-S (13.68 μg/m3). We also compare the usage of 4 statistical or machine learning models (Multiple Linear Regression, Random Forest, Gaussian Mixture Regression, and XGBoost) to correct low-cost sensors data, and find that XGBoost performs the best in testing (R2: 0.97, 0.94, 0.96; mean absolute error: 0.56, 0.80, and 0.68 μg/m3 for PurpleAir PA-II, Clarity Node-S, and Modulair-PM, respectively), but tree-based models do not perform well when correcting data outside the range of the colocation training. Therefore, we used Gaussian Mixture Regression to correct data from the network of 17 Clarity Node-S monitors deployed around Accra, Ghana, from 2018 to 2021. We find that the network daily average PM2.5 concentration in Accra is 23.4 μg/m3, which is 1.6 times the World Health Organization Daily PM2.5 guideline of 15 μg/m3. While this level is lower than those seen in some larger African cities (such as Kinshasa, Democratic Republic of the Congo), mitigation strategies should be developed soon to prevent further impairment to air quality as Accra, and Ghana as a whole, rapidly grow.

Humidity↗

Neural Network Development Tool (NETS)

Artificial neural networks formed from hundreds or thousands of simulated neurons, connected in manner similar to that in human brain. Such network models learning behavior. Using NETS involves translating problem to be solved into input/output pairs, designing network configuration, and training network. Written in C.

Baffes, Paul T.↗