Search NASASearch

SEARCH · Search NASA

Results for “data processing methods”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Evapotranspiration partitioning estimates from 8 methods from 47 NEON sites, 2019-2021

This dataset provides daily estimates of evapotranspiration (ET) and the transpiration-to-evapotranspiration ratio (T/ET) across 47 terrestrial National Ecological Observatory Network (NEON) sites spanning diverse environmental and biome conditions in the United States across three years of data (2019-2021). Daily ET is reported in both energy units (MJ m⁻² day⁻¹) and equivalent water depth (mm day⁻¹), assuming a constant latent heat of vaporization of 2.45 MJ/kg. The primary method uses a hybrid recurrent neural network–Penman–Monteith framework (RNN-PM), which integrates physically based surface energy balance constraints with data-driven learning to partition ET into transpiration and evaporation components. Model inputs include in situ meteorological observations (air temperature, vapor pressure deficit, wind speed, and radiation) combined with satellite-derived land surface temperature, leaf area index, and soil moisture. For benchmarking and uncertainty assessment, T/ET estimates from seven additional models are included: Priestley-Taylor Jet Propulsion Laboratory (PT-JPL), Penman-Monteith (P-M), Two-Source Energy Balance (TSEB), Support Vector Regression (SVR), and Categorical Boosting (CatBoost), among others—spanning empirical, machine-learning, and process-based approaches (see methods section or linked publication for detailed descriptions). Data Package Contents: The dataset a csv files containing daily ET and T/ET estimates for each site and model, along with associated metadata files these variables. Data can be accessed using common spreadsheet software (e.g., Microsoft Excel, LibreOffice) or programming environments such as R or Python. Together, these data support cross-site comparisons of ecosystem water use, evaluation of ET partitioning methods, and development of improved land–atmosphere exchange models.

EARTH SCIENCE > ATMOSPHERE

A Proxy Method to Bridge LCA Data Gaps Using Automated Material Classification and Probabilistic Under-Specification

Life cycle assessments (LCAs) are essential for understanding the environmental impacts of material production. However, gaps in life cycle inventory (LCI) data for material and chemical inputs present a key challenge for LCA practitioners, especially in the early design stages. Strategies for filling in these gaps require additional time and expertise, which can hinder the LCA’s completion. This study combined automatic material classification and probabilistic under-specification to create a time-efficient method to fill material LCI data gaps. To illustrate the proposed method, proxy environmental impact distributions were generated using publicly available material LCI data classified into the ChemOnt chemical taxonomy using the open-source chemical classification software ClassyFire. Input materials with data gaps were then classified into the same taxonomy, where proxy environmental impact values could be selected from the available distributions to quickly fill in any data gaps. Although these methods were applied to classify material production processes available in the Federal LCA Commons and Ecoinvent databases, they can be applied to any LCA database. This study shows that classifying materials by their chemical structure produces taxonomies with increased granularity relative to industrial classification, improving the ability of under-specified proxy data to be used for differentiating the environmental impacts of competing designs.

biological databases

Applying Gaussian Process Machine Learning and Modern Probabilistic Programming to Satellite Data to Infer CO 2 Emissions

Satellite data provides essential insights into the spatiotemporal distribution of CO 2 concentrations. However, many atmospheric inverse models fail to adequately incorporate the spatial and temporal correlations inherent in satellite observations and often lack rigorous methods for estimating parameters like spatial length scales. We introduce an inference model that processes the spatiotemporal covariance in satellite data and estimates hyperparameters such as covariance length scales. Our approach uses the Gaussian process (GP) machine learning (ML) and modern probabilistic programming languages (PPLs) to perform atmospheric inversions of emissions from satellite data. We develop a GP ML inversion system based on modern PPLs and the GEOS-Chem chemical transport model, simulating atmospheric CO 2 concentrations corresponding to the Orbiting Carbon Observatory-2/3 (OCO-2/3) data for July 2020. In our supervised learning framework, we treat the GEOS-Chem simulated data set as the target, with predictors derived by scaling the target with sector-specific factors hidden from the GP machine. Our results show that the GP model, combined with GPU-enabled PPLs, effectively retrieves true emission scaling factors and infers noise levels concealed within the data. This suggests that our method could be applied over larger areas with more complex covariance structures, enabling comprehensive analysis of the spatiotemporal patterns observed in OCO-2/3 and similar satellite data sets.

54 ENVIRONMENTAL SCIENCES

Uncertainty Quantification for Smooth Functional Data with Application to Material Properties

This document outlines a method for processing functional output (i.e., curves) for the ultimate purpose of sampling curves under specified input conditions for use in modeling and simulation uncertainty quantification (UQ) studies. A set of benchmark curves sufficiently representative of the relevant scenario(s) being simulated are provided to the process and formatted as described in Section 1. Principal Component Analysis (PCA) is utilized to discover the components of uncertainty in the benchmark curves and is outlined in Section 2. Section 3 describes the application of uncertainty quantification to the PCA results for the purpose of sampling curves to be used in UQ analysis. Section 4 applies these techniques to an example benchmark dataset. Concluding remarks are provided in the final section.

36 MATERIALS SCIENCE

Extracting the Breakout Distance from the ECOT Trajectories: Gaussian Process Regression Approach

Enhanced Corner Turning (ECOT) experiments provide an important metric of performance of high explosive (HE) formulations. The breakout distance is a single scalar value that characterizes the corner turning efficiency of an HE. Extracting the breakout distance from the raw ECOT results, whether experimental or simulated, is a conceptually straightforward procedure which, however, is non-unique, especially in the presence of noise. More specifically, this procedure involves numerical smoothing and selecting particular values for parameters of this smoothing introduces human bias. In this work, we propose to use the Gaussian process regression to analyze ECOT results. This analysis involves the effective smoothing of the data, thus allowing for accurate extraction of the breakout distance. Most importantly, the parameters of this smoothing can be inferred from the ECOT data itself, rendering the approach effectively parameter-free and thus diminishing the human bias. An additional benefit of the Gaussian process regression, being a statistical inference method, is that not just the value of the breakout distance, but also its confidence interval can be extracted from the data. This report introduces the Gaussian process regression, as applied to ECOT, and demonstrates its usefulness by extracting the breakout distances for a selection of experimental and simulated data.

45 MILITARY TECHNOLOGY, WEAPONRY, AND NATIONAL DEF

A Novel and Scalable Method for Microencapsulating Salt Hydrate Phase Change Materials in Core–Shell Fibers

Phase change materials (PCMs) are in high demand for applications such as thermal energy storage in buildings, electronics cooling, and thermal management of electric vehicle batteries and data centers. Among these materials, salt hydrate PCMs are particularly attractive due to their high thermal energy storage capacity and low cost. However, they suffer from two major issues: leakage in the melted phase and phase segregation during phase transitions. Microencapsulation is the primary process capable of addressing both of these challenges. However, there is no reliable or scalable method available for microencapsulating salt hydrate PCMs. As a result, the full potential of salt hydrates for building and data center applications has yet to be realized. In this work, we present an innovative method for the microencapsulation of salt hydrate PCMs using a co‐axial pushing technique. This process creates core–shell fibers, with the salt hydrate as the core and a polymer as the shell. Our approach demonstrates strong potential for scalable microencapsulation of salt hydrate PCMs. In conclusion, achieving scalability could enable their widespread use in applications such as data center cooling, battery thermal management, and building climate control.

Sharma, Jaswinder [Oak Ridge National Laboratory (

A physics informed bayesian optimization approach for material design: application to NiTi shape memory alloys

Abstract The design of materials and identification of optimal processing parameters constitute a complex and challenging task, necessitating efficient utilization of available data. Bayesian Optimization (BO) has gained popularity in materials design due to its ability to work with minimal data. However, many BO-based frameworks predominantly rely on statistical information, in the form of input-output data, and assume black-box objective functions. In practice, designers often possess knowledge of the underlying physical laws governing a material system, rendering the objective function not entirely black-box, as some information is partially observable. In this study, we propose a physics-informed BO approach that integrates physics-infused kernels to effectively leverage both statistical and physical information in the decision-making process. We demonstrate that this method significantly improves decision-making efficiency and enables more data-efficient BO. The applicability of this approach is showcased through the design of NiTi shape memory alloys, where the optimal processing parameters are identified to maximize the transformation temperature.

Chemistry

Recent Collaborations and Innovations to Demonstrate Next-Generation Techniques for Monitoring Subsurface Carbon Storage

World Carbon Capture, Utilization, and Storage (CCUS) Conference, Bergen, Norway, September 1–4, 2025. This talk provides a high-level overview of many novel and sustainable carbon storage-monitoring methods to accelerate the deployment of CCUS technologies at future CCUS sites across the United States. The Energy & Environmental Research Center’s work impacts the general CCUS industry by providing novel low-impact methods for tracking the injected plume’s migration and more autonomous data collection and processing techniques for performing assurance monitoring. Specifically, the results benefit 1) CCUS community members through knowledge sharing of lessons learned; 2) CCUS operators through commercialization of additional methods, including improvements to workflows and simplification of fieldwork; and 3) CCUS project stakeholders through implementation of low-impact and more autonomous monitoring solutions.

02 PETROLEUM

Multifidelity_Timeseries

SAND2025-03305O Multifidelity Timeseries is a user-friendly tool designed to create advanced models for analyzing time-series data. It offers three modeling options, allowing users to choose the best fit for their specific needs. The software efficiently processes multiple data sources without the need for complex sampling methods. It helps uncover patterns and insights using data. The result is it is easier to make informed decisions for projects. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Katona, Ryan [Sandia National Lab. (SNL-CA), Liver

The Integration and Mapping of an Open-Source National Well Resource to Inform Geologic Carbon Storage Site Selection and Risk Prevention: The CO2-Locate Database

Geologic carbon storage (GCS) offers a way to capture and permanently store CO₂ from fossil fuel operations in underground geologic structures, aiding in the transition to a carbon-neutral energy economy. However, CO₂ injection sites can experience gas leakage through existing wells that penetrate storage reservoirs, making knowledge of well locations and characteristics crucial for permitting, infrastructure reusability, and risk assessment in GCS. Currently, public wellbore data from state, federal, and tribal entities are inconsistent and fragmented, with gaps and redundancies. To address this, the National Energy Technology Laboratory (NETL) developed CO2-Locate, an open-source, geospatial database and online application. CO2-Locate integrates over 50 data sources from federal, state, and tribal entities, creating a standardized national well database. Funded by the Bipartisan Infrastructure Law, the database is publicly available through the Energy Data eXchange (EDX) and viewable via the CO2-Locate web mapping application. This tool allows users to query, filter, and visualize well data to support GCS planning, permitting, and risk assessments. This presentation covers the methods used to create CO2-Locate, including data acquisition, processing, attribute mapping, and integration, much of which is automated for future updates. The web mapping application and its role in GCS site selection will also be discussed.

Tetteh, Daniel A.

Digital Tools for the Preventive Conservation of Built Heritage: The Church of Santa Ana in Seville

Historic Building Information Modelling (HBIM) plays a pivotal role in heritage conservation endeavours, offering a robust framework for digitally documenting existing structures and supporting conservation practices. However, HBIM’s efficacy hinges upon the implementation of case-specific approaches to address the requirements and resources of each individual asset and context. This paper defines a flexible and generalisable workflow that encompasses various aspects (i.e., documentation, surveying, vulnerability assessment) to support risk-informed decision making in heritage management tailored to the peculiar conservation needs of the structure. This methodology includes an initial investigation covering historical data collection, metric and condition surveys and non-destructive testing. The second stage includes Finite Element Method (FEM) modelling and structural analysis. All data generated and processed are managed in a multi-purpose HBIM model. The methodology is tested on a relevant case study, namely, the church of Santa Ana in Seville, chosen for its historical significance, intricacy and susceptibility to seismic action. The defined level of detail of the HBIM model is sufficient to inform the structural analysis, being balanced by a more accurate representation of the alterations, through linked orthophotos and a comprehensive list of alphanumerical parameters. This ensures an adequate level of information, optimising the trade-off between model complexity, investigation time requirements, computational burden and reliability in the decision-making process. Field testing and FEM analysis provide valuable insight into the main sources of vulnerability in the building, including the connection between the tower and nave and the slenderness of the columns.

Chaves, Estefanía

Bayesian inference of anisotropic 2D small-angle scattering from sparse measurement

Here, we present a Bayesian inference framework for reconstructing anisotropic two-dimensional small-angle scattering (2D SAS) patterns from sparse, noisy, or partially missing data. The method combines a symmetry-aware angular basis with radial Gaussian process priors to enable accurate, training-free interpolation and denoising. Computational benchmarks demonstrate reliable recovery of both isotropic and high-order anisotropic features under severe data reduction. Experimental validations on stretched polymers, sheared wormlike micelles, and carbon fibers show improved fidelity and resolution compared to raw measurements, achieving comparable accuracy with up to 50-fold fewer detected neutrons. This approach enables quantitative structural analysis under low-flux, time-limited, or single-shot conditions, extending the applicability of 2D SAS techniques to compact neutron sources and mechanically driven soft matter systems undergoing transient structural changes.

Tung, Chi-Huan [Oak Ridge National Laboratory (ORN

Demonstration and Automation of Reflected Target Optical Measurement for Heliostats

Accurate optical surfaces are a primary driver of concentrated solar power plant performance. Errors in pointing and tracking mirrors, the canting of individual mirror facets, and the surface slope of the mirror itself can be caused by errors during assembly, transportation, wind loading, gravity, and many other sources. The tools that exist to measure these error sources today largely rely on fringe deflectometry (SOFAST, QDec, others), or photogrammetry with targets attached to the mirror surface. Since 2022, NREL has been developing a measurement method called the Reflected Target Non-intrusive Assessment (ReTNA) system. This system differs from most established methods in that we perform deflectometry with a pattern of coded targets, identified in space with photogrammetry. Reflected target systems have several advantages over traditional fringe deflectometry systems. Firstly, they can be operated in bright or ambient lighting, a challenge for fringe systems that use a projector and screen. Reflected target systems also can use a much lighter and less expensive target than projector-based systems. Lastly, 2D slope measurement can be solved from a single image, which leads to several advantages for accommodating faster measurements and smaller sized targets. These advantages make ReTNA particularly well-suited for applications where there are space or lighting constraints, like performing heliostat quality assurance on an assembly line. It's also useful when a lightweight, flexible system is needed, like for heliostat developers to quickly measure a new heliostat design at different orientations, to observe gravitational effects on the mirror surface shape. In the last year, significant improvements were made to this tool to make it more useful for these applications. These improvements were focused around validation of the ReTNA measurement system, and automation of the setup and measurement process. First, we present an improved ReTNA layout, for use on the heliostat assembly line. Next, we detail the various changes to the ReTNA software and computer vision methods to automate data collection in this new setup, and lessons learned from this process. The goal with this new setup is to perform a full heliostat surface characterization without removing the mirror from the assembly line. Lastly, we share results from several ReTNA validation studies undertaken over the last year. These include repeated ReTNA measurement on demonstration mirror facets, comparisons with other optical measurement tools, and some studies aimed at quantifying the uncertainty of ReTNA measurement under various constraints (mirror-target spacing, camera resolution, etc.). These results are compared with 2024 HelioCon performance targets, and our planned next steps for the ReTNA measurement system are presented.

CSP

Determining Magnetic Field Directions and Omni-directional Fluxes of Particles from STPSat-6 ZPS Plasma Measurements

This report summarizes our recent data derivation efforts on STPSat-6 ZPS plasma measurements, presenting final outcomes. We begin by outlining the methodology developed to determine local magnetic field directions from measured directional intensities of ~20 keV ions, leveraging symmetry in particles’ pitch-angle distributions, and the determined field directions over a storm period are shown with errors analyzed through comparison with NOAA GOES measurements. Next, we further assess the validity of the method and quantify its overall high performance while acknowledging certain caveats. Finally, given local magnetic field directions and ZPS data, we explain how omni-directional fluxes of particles can be derived from ZPS measurements with limited directional coverage, and estimate the related errors under several theoretical scenarios. The methods developed in this study can be applied for processing and augmenting ZPS data for the next, and the insights gained here can also inform instrument design in the future.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY

Autoregressive distributed lag-based dynamic uniformity modeling and monitoring approaches for superconductor manufacturing

High-temperature superconductors (HTS), known for their high efficiency and low energy loss, have found profound applications across various fields, driving the demand for long, uniformly performing tapes. However, ensuring uniform performance over extended lengths of HTS tapes, often characterized by the consistency of critical current, remains challenging due to fluctuations in growth conditions during manufacturing. Here, to elucidate the mechanisms underlying variations in tape uniformity and enable real-time monitoring of associated parameters, we propose an Autoregressive Distributed Lag (ADL)-based Dynamic Uniformity Modeling and Monitoring (ADUM2) approach. This method integrates uniformity measurement, the identification of critical process parameters and real-time monitoring within the manufacturing process. The ADUM2 approach is applied to the advanced metal organic chemical vapor deposition (A-MOCVD) process, a pilot-scale method for superconductor manufacturing. Our model demonstrates superior performance compared to benchmark methods, accounting for over 80% of the total variance in the data and identifying 13 key process parameters influencing the uniformity of HTS tapes. This study offers significant insights into the high-temperature superconductor manufacturing process and holds the potential to facilitate the production of cost-effective, uniformly performing long superconducting tapes in the future.

autoregressive distributed lag analysis

MARVEL Reactor Digital Engineering Developments

The MARVEL reactor project has served to introduce a new generation of engineers to the processes required to transform a reactor design from simply an idea on paper into what will be an approved, constructed, and operational nuclear power system. Much as there have been advances in materials, analysis, and evaluation methodologies over the 50 years since the last reactor was built at INL, so too has the technology for managing the engineering process itself advanced. Digital Engineering tools and methods provide improved coordination between previously siloed engineering disciplines, reduced burdens of non-value-added data transcription processes and bring forward insights and improvements that might otherwise fall later in the design stage, where changes are much more costly. While the tools and techniques to support the full digital engineering vision are not yet complete, the MARVEL design processes provide valuable demonstrations and validations of key aspects and illuminate further areas for implementation by subsequent projects.

21 - SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLAN

High-Throughput Microstructural Characterization and Process Correlation Using Automated Electron Backscatter Diffraction

The need to optimize the processing conditions of additively manufactured (AM) metals and alloys has driven advances in throughput capabilities for material property measurements such as tensile strength or hardness. High-throughput (HT) characterization of AM metal microstructure has fallen significantly behind the pace of property measurements due to intrinsic bottlenecks associated with the artisan and labor-intensive preparation methods required to produce highly polished surfaces. This inequality in data throughput has led to a reliance on heuristics to connect process to structure or structure to properties for AM structural materials. In this study, we show a transformative approach to achieve laser powder bed fusion (LPBF) printing, HT preparation using dry electropolishing and HT electron backscatter diffraction (EBSD). This approach was used to construct a library of > 600 experimental EBSD sample sets spanning a diverse range of LPBF process conditions for AM Kovar. This vast library is far more expansive in parameter space than most state-of-the-art studies, yet it required only approximately 10 labor hours to acquire. Build geometries, surface preparation methods, and microscopy details, as well as the entire library of >600 EBSD data sets over the two sample design versions, have been shared with intent for the materials community to leverage the data and further advance the approach. Using this library, we investigated process–structure relationships and uncovered an unexpected, strong dependence of microstructure on location within the build, when varied, using otherwise identical laser parameters.

Characterization and Analytical Technique

Constraining Galaxy-Halo connection using machine learning

We investigate the potential of machine learning (ML) methods to model small-scale galaxy clustering for constraining Halo Occupation Distribution (HOD) parameters. Our analysis reveals that while many ML algorithms report good statistical fits, they often yield likelihood contours that are significantly biased in both mean values and variances relative to the true model parameters. This highlights the importance of careful data processing and algorithm selection in ML applications for galaxy clustering, as even seemingly robust methods can lead to biased results if not applied correctly. ML tools offer a promising approach to exploring the HOD parameter space with significantly reduced computational costs compared to traditional brute-force methods if their robustness is established. Using our ANN-based pipeline, we successfully recreate some standard results from recent literature. Properly restricting the HOD parameter space, transforming the training data, and carefully selecting ML algorithms are essential for achieving unbiased and robust predictions. Among the methods tested, artificial neural networks (ANNs) outperform random forests (RF) and ridge regression in predicting clustering statistics, when the HOD prior space is appropriately restricted. We demonstrate these findings using the projected two-point correlation function (w p (r p )), angular multipoles of the correlation function (ξ ℓ (r)), and the void probability function (VPF) of Luminous Red Galaxies from Dark Energy Spectroscopic Instrument mocks. Our results show that while combining w p (r p ) and VPF improves parameter constraints, adding the multipoles ξ 0 , ξ 2 , and ξ 4 to w p (r p ) does not significantly improve the constraints.

cosmology