Search NASA⌕ Search

SEARCH · Search NASA

Results for “representation learning”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 469 records · Page 26

LossLens: Diagnostics for Machine Learning Through Loss Landscape Visual Analytics

Modern machine learning often relies on optimizing a neural network's parameters using a loss function to learn complex features. Beyond training, examining the loss function with respect to a network's parameters (i.e., as a loss landscape) can reveal insights into the architecture and learning process. While the local structure of the loss landscape surrounding an individual solution can be characterized using a variety of approaches, the global structure of a loss landscape, which includes potentially many local minima corresponding to different solutions, remains far more difficult to conceptualize and visualize. To address this difficulty, we introduce LossLens, a visual analytics framework that explores loss landscapes at multiple scales. LossLens integrates metrics from global and local scales into a comprehensive visual representation, enhancing model diagnostics. Here we demonstrate LossLens through two case studies: visualizing how residual connections influence a ResNet-20, and visualizing how physical parameters influence a physics-informed neural network (PINN) solving a simple convection problem.

97 MATHEMATICS AND COMPUTING↗

Neural Predictors of Visuomotor Adaptation Rate and Multi-Day Savings

Recent studies of sensorimotor adaptation have found that individual differences in task-based functional brain activation are associated with the rate of adaptation and savings at subsequent sessions. However, few studies to date have investigated offline neural predictors of adaptation and multi-day savings. In the present study, we explore whether individual differences in the rate of visuomotor adaptation and multi-day savings are associated with differences in resting state functional connectivity and gray matter volume. Thirty-four participants performed a manual adaptation task during two separate test sessions, on average 9 days apart. We found that resting state functional connectivity strength between sensorimotor, anterior cingulate, and temporoparietal areas of the brain was a significant predictor of adaptation rate during the early, cognitive phase of practice. In contrast, default mode network functional connectivity strength was found to predict late adaptation rate and savings on day two, which suggests that these behaviors may rely on overlapping processes. We also found that gray matter volume in temporoparietal and occipital regions was a significant predictor of early learning, whereas gray matter volume in superior posterior regions of the cerebellum was a significant predictor of late adaptation. The results from this study suggest that offline neural predictors of early adaptation facilitate the cognitive mechanisms of sensorimotor adaptation, with support from by the involvement of temporoparietal and cingulate networks. In contrast, the neural predictors of late adaptation and savings, including the default mode network and the cerebellum, likely support the storage and modification of newly acquired sensorimotor representations. These findings provide novel insights into the neural processes associated with individual differences in sensorimotor adaptation.

Cassady, Kaitlin↗

Advancing the Representation of Human Actions in Large‐Scale Hydrological Models: Challenges and Future Research Directions

Characterizing the impact of human actions on terrestrial water fluxes and storages at multi-basin, continental, and global scales has long been on the agenda of scientists engaged in climate science, hydrology, and water resources systems analysis. This need has resulted in a variety of modeling efforts focused on the representation of water infrastructure operations. Yet, the representation of human-water interactions in large-scale hydrological models is still relatively crude, fragmented across models, and often achieved at coarse resolutions (~10–100 km) that cannot capture local water management decisions. In this commentary, we argue that the concomitance of four drivers and innovations is poised to change the status quo: “hyper-resolution” hydrological models (~0.1–1 km), multi-sector modeling, satellite missions able to monitor the outcome of human actions, and machine learning are creating a fertile environment for human-water research to flourish. We then outline four challenges that chart future research in hydrological modeling: (a) creating hyper-resolution global data sets of water management practices, (b) improving the characterization of anthropogenic interventions on water quantity, stream temperature, and sediment transport, (c) improving model calibration and diagnostic evaluation, and (d) reducing the computational requirements associated with the successful exploration of these challenges. Overcoming them will require addressing modeling, computational, and data development needs that cut across the hydrology community, thereby requiring a major communal effort.

catchment hydrology↗

Program Helps In Analysis Of Failures

Failure Environment Analysis Tool (FEAT) computer program developed to enable people to see and better understand effects of failures in system. User selects failures from either engineering schematic diagrams or digraph-model graphics, and effects or potential causes of failures highlighted in color on same schematic-diagram or digraph representation. Uses digraph models to answer two questions: What will happen to system if set of failure events occurs? and What are possible causes of set of selected failures? Helps design reviewers understand exactly what redundancies built into system and where there is need to protect weak parts of system or remove them by redesign. Program also useful in operations, where it helps identify causes of failure after they occur. FEAT reduces costs of evaluation of designs, training, and learning how failures propagate through system. Written using Macintosh Programmers Workshop C v3.1. Can be linked with CLIPS 5.0 (MSC-21927, available from COSMIC).

Stevenson, R. W.↗

Reanalysis Activities at the NASA Global Modeling and Assimilation Office

This talk presents an overview of recent reanalysis activities at the NASA Global Modeling and Assimilation Office (GMAO) as part of a multi-faceted strategy towards an Integrated Earth System retrospective analysis, coupling components of the atmosphere, ocean, chemistry, land, and ice. While elements of the atmosphere-ocean coupled Goddard Earth Observing System (GEOS) model and data assimilation are being actively developed, a suite of reanalysis products is designed to provide further understanding of key aspects of Earth system coupling in a reanalysis context: The baseline atmospheric reanalysis, the GEOS Retrospective analysis for the early 21st Century (GEOS-R21C), features recent advances in the GEOS model and data assimilation, and targets the NASA’s Earth Observing System EOS and post-EOS satellite observations; GEOS-IT, a user-tailored low-resolution atmospheric reanalysis, serves as a second baseline to the NASA Instrument Teams for validation and calibration and drives a one-way coupled ocean reanalysis, GEOSIT-Ocean; PolarMERRA, a high-resolution downscaled product for the polar regions, focuses on improving the representation of polar atmospheric processes with an assessment of current cryospheric biases, and targeted improvements to surface sea ice and glacier conditions; Finally, R21C-Chem, an off-line atmospheric chemistry and composition reanalysis, includes both tropospheric and stratospheric trace gases. The diversity of these reanalysis activities presents unique opportunities for collaborations cross-teams/institutions, with new commercial data partners, and with end-user groups. This talk will discuss these opportunities and explore leveraging the lessons learned along the way on key drivers in Earth system interactions as we converge towards the next generation of the Modern-Era Retrospective analysis for Research and Applications (MERRA) suite.

Amal El Akkraoui↗

Geothermal Power Systems Analysis: Outcome of Industry Stakeholders Workshop: Preprint

Geothermal cost and performance evaluation implemented via technoeconomic assessment (TEA) modeling is critical for the Department of Energy (DOE) and other geothermal industry stakeholders in assessing the current state of geothermal technologies and to identify existing hurdles to commercially viable geothermal development. The Geothermal Electricity Technology Evaluation Model (GETEM) is a major TEA tool used in estimating the economic feasibility and levelized cost of energy (LCOE) of conventional hydrothermal systems and enhanced geothermal systems (EGS). Since 2021, GETEM has been transitioning from an intricate spreadsheet model to a user-friendly tool within the System Advisor Model (SAM) developed by the National Renewable Energy Laboratory (NREL). Apart from enabling an expanded visibility of the geothermal model among other renewable resources, having GETEM in SAM has the advantage of simulation automation, better usability, updates tracking, active user inputs/feedback, and extended financial modeling. GETEM is used in developing supply curves for the Annual Technology Baseline (ATB). The ATB data are inputs to the Renewable Energy Potential (reV) and the Regional Energy Deployment System (ReEDS) models. The geothermal module in NREL’s reV model assesses the geothermal energy potential in the conterminous United States by defining the geospatial intersection of geothermal resources with existing grid infrastructure within the constraint of land use characteristics. The ReEDS model is a capacity expansion model used for simulating the long-term build-out and operation of the US generation and transmission system based on current energy costs and policies. To ensure enhanced representation of current industry trends in our model transitions and development, we organized a two-day virtual workshop to elicit geothermal industry stakeholder input and recommendations on our current approaches and assumptions on technoeconomic, resource assessment, and deployment scenarios modeling of geothermal technologies. Participants included developers, operators, investors, regulatory agencies, system modelers, national laboratory researchers, consultants, and other stakeholders. In this workshop, we gained stakeholder insights on current geothermal plant performance (i.e., capacity factors), updated drilling costs and learning curves, and next generation technologies such as closed loop and superhot rock geothermal. Other outcomes from this workshop and its impact on future geothermal development feasibility, resource availability, and capacity expansion studies are compiled and discussed.

Annual Technology Baseline↗

ReLU, Sparseness, and the Encoding of Optic Flow in Neural Networks

Accurate self-motion estimation is critical for various navigational tasks in mobile robotics. Optic flow provides a means to estimate self-motion using a camera sensor and is particularly valuable in GPS- and radio-denied environments. The present study investigates the influence of different activation functions—ReLU, leaky ReLU, GELU, and Mish—on the accuracy, robustness, and encoding properties of convolutional neural networks (CNNs) and multi-layer perceptrons (MLPs) trained to estimate self-motion from optic flow. Our results demonstrate that networks with ReLU and leaky ReLU activation functions not only achieved superior accuracy in self-motion estimation from novel optic flow patterns but also exhibited greater robustness under challenging conditions. The advantages offered by ReLU and leaky ReLU may stem from their ability to induce sparser representations than GELU and Mish do. Our work characterizes the encoding of optic flow in neural networks and highlights how the sparseness induced by ReLU may enhance robust and accurate self-motion estimation from optic flow.

97 MATHEMATICS AND COMPUTING↗

Robust Spectral Anomaly Detection in EELS Spectral Images via 3D Convolutional Variational Autoencoders

Abstract A 3D Convolutional Variational Autoencoder (3D‐CVAE) is introduced for automated anomaly detection in electron energy‐loss spectroscopy spectrum imaging (EELS‐SI) data. This approach leverages the full 3D structure of EELS‐SI data to detect subtle spectral anomalies while preserving both spatial and spectral correlations across the datacube. By employing cross‐entropy loss and training on bulk spectra, the model learns to reconstruct bulk features characteristic of the defect‐free material. In exploring methods for anomaly detection, both the 3D‐CVAE approach and principal component analysis (PCA) are evaluated, testing their performance using FeL‐edge ΔEpeak shifts designed to simulate material defects. These results show that 3D‐CVAE achieves superior anomaly detection and maintains consistent performance across various shift magnitudes. The method demonstrates clear bimodal separation between bulk and anomalous spectra, enabling reliable classification. Further analysis verifies that lower‐dimensional representations are robust to anomalies in the data. While performance advantages over PCA diminish with decreasing anomaly concentration, our method maintains high reconstruction quality even in challenging, noise‐dominated spectral regions. This approach provides a robust framework for unsupervised automated detection of spectral anomalies in EELS‐SI data, particularly valuable for analyzing complex material systems.

Chemistry↗

Mapping wall-to-wall fractional cover of Arctic tundra plant functional types in Alaska using 20-m spatial resolution satellite imagery and harmonized plot observations

Estimates of fractional cover (fCover) across given land surfaces are used to assess, and often model, vegetation composition and diversity, which are crucial for understanding the health and functioning of terrestrial ecosystems. Remote sensing provides a useful means for scaling local, plot-measured fCover estimates to regional scales. Leveraging a recently synthesized and harmonized plot database, this study generated wall-to-wall maps of fCover for six Alaskan-Arctic plant functional types (PFT), including non-vascular plants, forbs, graminoids, and deciduous and evergreen shrubs, using 20-m satellite data (Sentinel-1, Sentinel-2, ArcticDEM) using a machine learning regression approach, specifically the random forest (RF) algorithm, which is well-suited for handling nonlinear relationships and high-dimensional satellite datasets. This study additionally addressed the spatio-temporal inconsistencies e.g., sampling scale, plot size, and collection year in plot measured fCover by adopting a multivariate outlier detection approach—Cook’s distance—to identify high-quality plots for model training and validation. Our approach achieves high accuracy (R 2 = 0.59–0.93, root mean squared errors = 0.02–0.10 for all PFTs) between plot-observed and satellite-derived fCover when using high-quality plot samples. The mapped fCover characterizes the spatial patterns of different PFTs across the tundra biome at a 20-m resolution, providing key information needed for improved representation of Arctic tundra vegetation in terrestrial biosphere models to better understand climate-vegetation feedback across the Arctic tundra.

Arctic tundra↗

A Comprehensive Machine Learning Model for Metal–Ligand Binding Prediction: Applications in Chemistry and Biology

A machine-learning (ML) model that predicts metal–ligand binding constants was developed using the open-source Chemprop software. The model was trained on over 30,000 experimental log K 1 values, which include both protonation and metal–ligand stability constants, comprising over 3500 ligands and 10 2 metal ions from 73 total elements, thus generalizing beyond existing limited approaches, which focus only on specific metals or ligand families. The best-performing model included a combination of SMILES-based molecular representations along with descriptors for the metal ion and experimental conditions. It had an external test R 2 value of 0.942, and MAE value of 0.834. A “SMILES-only” simpler version also produced accurate predictions and preserved the binding trends, serving as a quick and easily accessible alternative for users without computational expertise. The SMILES-only model performed comparably to density functional theory (DFT) calculations but utilized a fraction of the computational resources. The model was successfully applied across diverse domains, including bioinorganic chemistry, heavy metal remediation, and sensor development and demonstrated its effectiveness as a rapid and reliable screening tool for both academic and industrial uses.

Ligands↗

Integrated System Planning: Emerging Software Requirements in the Power Industry

Power system planning software remains fragmented across organizational boundaries, with specialized tools for capacity expansion, production cost modeling, power flow, and dynamic analysis operating on incompatible data models and assumptions. This article argues that the fragmentation is not merely a technical problem but a predictable consequence of Conway's law: software architectures mirror the departmental structures within which they are developed. Regulatory milestones like Federal Energy Regulatory Commission (FERC) Order 888 formalized these divisions, but the roots trace back to the distinct engineering disciplines-mechanical, chemical, and electrical-that staffed generation and transmission planning departments in vertically integrated utilities. As the industry moves toward integrated system planning (ISP) that coordinates generation, transmission, and distribution investment decisions, the software ecosystem must evolve accordingly. We identify five categories of software requirements to enable this transition: coherent data inputs decoupled from individual applications, unified and extensible data schemas, modular component representations that support multiple abstraction levels, lifecycle management of planning datasets, and well-defined application programming interface (API) contracts that separate data exchange from algorithmic control. We examine how these requirements interact with three common workflow patterns-serial gate clearing, sequential multiapplication, and convergence oriented-and discuss the interface design principles each demands. We then outline a vision for platform-based planning architectures where specialized analytical services compose through standardized interfaces and where artificial intelligence (AI)/machine learning (ML) tools augment decision support within a disciplined software infrastructure. The practices proposed here offer a path from today's siloed tool collections toward collaborative planning ecosystems capable of handling the complexity of modern power system transformation.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Principles for Architecting Autonomous Systems

This paper distills principles for developing autonomous systems based on experience and lessons learned from past efforts. The purpose of these principles is to establish a common understanding and knowledge of architectural elements to guide the development of next-generation multi-mission autonomous systems and ensure the safe and productive operation of space assets. An attempt has been made to ground these principles in fundamentals that should withstand the test of time while allowing for and enabling the advancement of technologies. They are not intended to prescribe a design nor a software representation. There may be multiple designs that can honor these principles. These principles are focused on autonomy for robotic assets. As such, they do not address autonomy for crewed assets nor autonomy that can collectively generate intelligent behavior without top-level system cognizance (e.g., intelligent swarm behavior). These areas would be a subject of future efforts.

Day, John↗

Enhancing Long-Term Trend Simulation of OH Through the Synergy of Model Simulations and Aura Ozone Monitoring Instrument (OMI) NO 2 and HCHO Retrievals

During the last few years, tremendous progress has been made to develop an efficient parameterization module using agile machine learning techniques. The aim of this module is to provide dynamic response of the tropospheric hydroxyl radical (OH) to its major drivers, including trace gases, aerosols, clouds, and meteorology. This module, named ECCOH (pronounced “echo”) and implemented in NASA’s GEOS-5 global model, offers an unrealized opportunity to unravel the convoluted response of OH to its underlying drivers while approaching the accuracy of full-chemistry without incurring excessive computational costs, making it suitable for climate models. However, the accurate representation of OH in ECCOH poses challenges due to the lack of representation of some of its critical inputs such as the abundance of NO 2 and HCHO concentrations. As such, we leverage the well-characterized satellite observations of NO2 and HCHO columns from Aura OMI to enhance their representation in ECCOH using an optimal interpolation method for the time period of 2005 - present. We show how the inclusion of OMI information can affect the spatiotemporal variability and long-term trends of OH, CO, and CH 4 across the globe. Additionally, we underscore the necessity of obtaining high-fidelity information regarding tropospheric ozone from the southern hemisphere from space, a region currently lacking full verification in models, posing a challenge to get a reasonable amount of chemical sink for CH 4 .

OH↗

Document Classification Techniques for Aviation Letters of Agreement

Often when working with technical documents, it is helpful to classify them into specific categories. In this paper, we conduct a thorough review of natural language processing techniques to perform this classification task on Letters of Agreement (LOAs), technical aviation documents outlining rules for utilizing US airspace. We evaluate multiple techniques, including Transfer Learning, for representing the text in the documents as embeddings: unigram and bigram Term Frequency Inverse Document Frequency (TFIDF), Word2Vec, Doc2Vec, GloVe and RoBERTa. We investigate a wide range of classification models: K-Nearest Neighbors, Random Forest, Support Vector Machines (SVM), Logistic Regression, Naive Bayes, Feed-Forward Neural Network, Convolutional Neural Networks (CNNs) and Long-Short Term Memory (LSTM). By comparing the different methods, we found the best overall approach for our task was to use unigram TFIDF representations with SVM while also gaining insight into how the other methodologies performed on a small technical datasets.

Aayushi Batra↗

Geothermal Power Systems Analysis: Outcome of Industry Stakeholders Workshop

Geothermal cost and performance evaluation implemented via techno-economic assessment (TEA) modeling is critical for the U.S. Department of Energy (DOE) and other geothermal industry stakeholders in assessing the current state of geothermal technologies and to identify existing hurdles to commercially viable geothermal development. The Geothermal Electricity Technology Evaluation Model (GETEM) is a major TEA tool used in estimating the economic feasibility and levelized cost of energy (LCOE) of conventional hydrothermal systems and enhanced geothermal systems (EGS). Since 2021, GETEM has been transitioning from an intricate spreadsheet model to a user-friendly tool within the System Advisor Model (SAM) developed by the National Renewable Energy Laboratory (NREL). Apart from enabling an expanded visibility of the geothermal model among other renewable resources, having GETEM in SAM has the advantage of simulation automation, better usability, updates tracking, active user inputs/feedback, and extended financial modeling. GETEM is used in developing supply curves for NREL's Annual Technology Baseline (ATB), which provides inputs to the Renewable Energy Potential (reV) and the Regional Energy Deployment System (ReEDS) models. The geothermal module in NREL's reV model assesses the geothermal energy potential in the conterminous United States by defining the geospatial intersection of geothermal resources with existing grid infrastructure within the constraint of land use characteristics. The ReEDS model is a capacity expansion model used for simulating the long-term build-out and operation of the U.S. generation and transmission system based on current energy costs and policies. To ensure enhanced representation of current industry trends in our model transitions and development, we organized a two-day virtual workshop to elicit geothermal industry stakeholder input and recommendations on our current approaches and assumptions on techno-economic, resource assessment, and deployment scenarios modeling of geothermal technologies. Participants included developers, operators, investors, regulatory agencies, system modelers, national laboratory researchers, consultants, and other stakeholders. In this workshop, we gained stakeholder insights on current geothermal plant performance (i.e., capacity factors), updated drilling costs and learning curves, and next-generation technologies such as closed-loop and superhot rock geothermal. Other outcomes from this workshop and its impact on future geothermal development feasibility, resource availability, and capacity expansion studies are compiled and discussed.

annual technology baseline↗

Scalable 3D reconstruction for X-ray single particle imaging with online machine learning

X-ray free-electron lasers offer unique capabilities for measuring the structure and dynamics of biomolecules, helping us understand the basic building blocks of life. Notably, high-repetition-rate free-electron lasers enable single particle imaging, where individual, weakly scattering biomolecules are imaged under near-physiological conditions with the opportunity to access fleeting states that cannot be captured in cryogenic or crystallized conditions. Existing X-ray single particle reconstruction algorithms, which estimate the particle orientation for each image independently, are slow and memory-intensive when handling the massive datasets generated by emerging free-electron lasers. Here, we introduce X-RAI (X-Ray single particle imaging with Amortized Inference), an online reconstruction framework that estimates the structure of 3D macromolecules from large X-ray single particle datasets. X-RAI consists of a convolutional encoder, which amortizes pose estimation over large datasets, as well as a physics-based decoder, which employs an implicit neural representation to enable high-quality 3D reconstruction in an end-to-end, self-supervised manner. We demonstrate that X-RAI achieves state-of-the-art performance for small-scale datasets in simulation and challenging experimental settings and demonstrate its unprecedented ability to process large datasets containing millions of diffraction images in an online fashion. These abilities signify a paradigm shift in X-ray single particle imaging towards real-time reconstruction.

Computer science↗

AIVT: Inference of turbulent thermal convection from measured 3D velocity data by physics-informed Kolmogorov-Arnold networks

We propose the artificial intelligence velocimetry-thermometry (AIVT) method to reconstruct a continuous and differentiable representation of the temperature and velocity in turbulent convection from measured three-dimensional (3D) velocity data. AIVT is based on physics-informed Kolmogorov-Arnold networks and trained by optimizing a loss function that minimizes residuals of the velocity data, boundary conditions, and governing equations. We apply AIVT to a set of simultaneously measured 3D temperature and velocity data of Rayleigh-Bénard convection, obtained by combining particle image thermometry and Lagrangian particle tracking. This enables us to directly compare machine learning results to true volumetric, simultaneous temperature and velocity measurements. We demonstrate that AIVT can reconstruct and infer continuous, instantaneous velocity and temperature fields and their gradients from sparse experimental data at a high resolution, providing an additional approach for understanding thermal turbulence.

Science & Technology - Other Topics↗

Advancing AI-Driven Analysis in X-ray Absorption Spectroscopy: Spectral Domain Mapping and Universal Models

In recent years, rapid progress has been made in developing artificial intelligence (AI) and machine learning (ML) methods for X-ray absorption spectroscopy (XAS) analysis. Compared to traditional XAS analysis methods, AI/ML approaches offer dramatic improvements in efficiency and help eliminate human bias. To advance this field, we advocate an AI-driven XAS analysis pipeline that features several interconnected key building blocks: benchmarks, workflows, databases, and AI/ML models. Specifically, we present two case studies for XAS ML. In the first study, we demonstrate the importance of reconciling the discrepancies between simulation and experiment using spectral domain mapping (SDM). Our ML model, which is trained solely on simulated spectra, predicts an incorrect oxidation state trend for Ti atoms in a combinatorial zinc titanate film. After transforming the experimental spectra into a simulation-like representation using SDM, the same model successfully recovers the correct oxidation state trend. In the second study, we explore the development of universal XAS ML models that are trained on the entire periodic table, which enables them to leverage common trends across elements. Looking ahead, we envision that an AI-driven pipeline can unlock the potential of real-time XAS analysis to accelerate scientific discovery.

36 MATERIALS SCIENCE↗