Search NASA⌕ Search

SEARCH · Search NASA

Results for “Data Science Model”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 343 records · Page 19

Novel Design and Fabrication of a High Frequency Transient Heat Flux Sensor for Use in an RDE

Rotating detonation engine (RDE) combustion systems have been a topic of interest in the pressure gain combustion community for their benefits over traditional gas turbine engine combustors. However, cooling requirements for these engines are significantly higher and less predictable than non-detonating engines. To understand the high-speed heat transfer dynamics inside an RDE, a novel, high-frequency heat flux gage is presented. This study aims to design a robust, single-sided sensor that can withstand the high temperature and harsh environment of an RDE for extended durations. Sensor bench testing is performed using a hot plate as a heat source, and the sensor response is compared to a finite-element analysis (FEA) model. The sensor response is then tested inside a water-cooled RDE and the wall heat flux is compared to calorimetry data.

rotating detonation engines↗

Alaska Meteorology, Energy, and Transmission (MET) Toolkit

The Alaska MET (Meteorology, Energy, and Transmission) Toolkit is the National Laboratory of the Rockies' (NLR) new flagship atmospheric dataset, designed to support comprehensive long-term planning and operations across the entire power sector. Serving as the regional counterpart to CONUS-wide HRRR MET Toolkit, this dataset provides a comprehensive, high-fidelity meteorological record covering Alaska.The Alaska MET Toolkit is delivered at an hourly resolution on a standardized 2-km horizontal grid. This dataset is repackaged from the National Oceanic and Atmospheric Administration's (NOAA) operational High-Resolution Rapid Refresh for Alaska (HRRR-AK) forecasts. Spanning from 2019 to 2025, it overcomes the technical barriers of native weather models by providing spatial regridding from the native 3-km HRRR-AK horizontal resolution to a 2-km grid, temporal gap-filling, and vertical interpolation at key energy-relevant heights. By delivering highly accurate, validation-backed data across a comprehensive suite of atmospheric variables - including temperature, pressure, humidity, and wind characteristics - the Alaska MET Toolkit provides a highly accessible and strictly standardized foundation for modern power system modeling.

17 WIND ENERGY↗

Hector V3.2.0: functionality and performance of a reduced-complexity climate model

Abstract. Hector is an open-source reduced-complexity climate–carbon cycle model that models critical Earth system processes on a global and annual basis. Here, we present an updated version of the model, Hector V3.2.0 (hereafter Hector V3), and document its new features, implementation of new science, and performance. Significant new features include permafrost thaw, a reworked energy balance submodel, and updated parameterizations throughout. Hector V3 results are in good general agreement with historical observations of atmospheric CO2 concentrations and global mean surface temperature, and the future temperature projections from Hector V3 are consistent with more complex Earth system model output data from the sixth phase of the Coupled Model Intercomparison Project. We show that Hector V3 is a flexible, performant, robust, and fully open-source simulator of global climate changes. We also note its limitations and discuss future areas for improvement and research with respect to the model's scientific, stakeholder, and educational priorities.

54 ENVIRONMENTAL SCIENCES↗

Data and Scripts associated with a manuscript on ecosystem responses to wildfires in the Columbia River Basin

This data package is associated with the publication “Ecosystem leaf area, gross primary production, and evapotranspiration responses to wildfire in the Columbia River Basin” submitted to Biogeosciences (Shi et al., 2024; doi: 10.22541/au.171053013.30286044/v1). In this research, data products, leaf area index (LAI), gross primary production (GPP), and evapotranspiration (ET), from the Moderate Resolution Imaging Spectroradiometer (MODIS) are used to quantify the resistance and resilience of different ecosystem types in the Columbia River Basin (CRB). A machine learning algorithm, random forest (RF), was used to examine the impacts of precipitation, vapor pressure deficit (VPD), and burn severity from Monitoring Trends in Burn Severity (MTBS) on ecosystem resilience. The data package includes the processed MODIS data products, precipitation, VPD, and burn severity in 138 fire regions in CRB and the input files for RF model training. This data package includes six folders. The MODIS products are included in three MODIS_* folders with shell scripts for data clipping and *ncl files for data processing: (1) “/MODIS_LAI_CRB”; (2) “/MODIS_GPP_CRB”; and (3) “/MODIS_ET_CRB”. All the processed data for each fire event are NetCDF formatted. The MTBS burn severity data and the shell and *ncl scripts used for data processing are in the folder named (4) “MTBS_fire”. The ERA meteorological fields and the data processing scritps are in (5) “ERA_Var_CR”. All the scripts for figure development are in the format of *ncl and in the folder (6) “paper_scripts”. See the file ending in “flmd.csv” for a list of all files contained in this data package and descriptions for each. Tabular column headers and units are described in the data dictionary file ending in “dd.csv”.

54 ENVIRONMENTAL SCIENCES↗

Videos, photos, and AI-derived grain size data associated with “High-throughput AI Video Surveys Enable Reproducible Multiscale Sediment Size Mapping, with Implications for Hydrobiogeochemical Parameterization”

NOTE: The manuscript associated with this data package is currently in review. The data may be revised based on reviewer feedback. Upon manuscript acceptance, this data package will be updated with the final dataset and additional metadata. This data package is associated with the manuscript “High-throughput AI Video Surveys Enable Reproducible Multiscale Sediment Size Mapping, with Implications for Hydrobiogeochemical Parameterization” under review. This data package includes five data types: 1) raw photos and videos from drone survey and walking smartphone surveys; 2) images derived from raw videos; 3) manual labeling of reference scales; 4) metadata for all images and photo resolution derived from artificial intelligence (AI) models or manual labels, 5) grain size data obtained from AI models for all photos, 6) metadata and grain size data after quality control, 7) summaries of sample efficiency for all data, and 8) computational fluid dynamics (CFD) data used to support hydro-biogeochemical (HBGC) parameter estimation. Such data is used to 1) demonstrate significant improvements in accuracy, efficiency, and quality control for grain size data collection with the help of AI models, 2) study the spatial heterogeneity of grain size and observation reproducibility based on tens of thousands of data points generated by the AI models, and 3) evaluate the impacts of grain size heterogeneity on key HBGC parameters across sediment-to-reach and hourly-to-yearly scales. In particular, the data package contains 116 folders and 179696 files. The files include 41 videos in .mov format, 64047 photos in .jpg format, 13541 video-derived photos in .png format, 12747 segmentation mask data in .tif format, 12747 segmentation data in .json format, 24771 .csv files that with metadata and grain size for each individual photo as well as water depth and velocity data from CFD and observation, 51791 .txt files of raw AI predicted labels, and 11 flight record data in .srt format. The summary for all metadata and grain size statistics information is included in “Scales_V3_NG.csv” and “Statistics_V3_NG.csv”. The summary for data that pass data quality control (QC) level 0-2 is included in “QCStatistics_V3_NG.csv”. The QC level 0 represents photos whose photo resolution is positive, excluding photos that miss reference scale. The QC level 1 means reference scale circularity uncertainty is less than 5% for smartphone images while representing photo resolution is larger than 0.44 mm/pixel for drone images. The QC level 2 means excluding photos whose grain number is less than 100, a minimum number of grains recommended by classic literature. The summary for each video’s name, length, frame rates, survey area, grain number, survey efficiency, etc. can be found in “QCSummary_V3_NG.csv”. The summary for site name, GPS coordinates, and number of images at each site can be found in “SitesSummary_V3_*.csv” files. Overall computational efficiency summary is reported in Table 4 of accompanying manuscript. Additionally, the nitrate concentration data used in this work was downloaded from an existing dataset published on ESS-DIVE (Boat-Dragged Sensor Hanford Reach.csv; Conner A. et al., 2020). We thank the United States Forest Service, Washington Department of Fish and Wildlife, Washington Department of Natural Resources, Cowiche Canyon Conservatory, Port of Benton, and the Confederated Tribes and Bands of the Yakama Nation for access to field locations where the data were collected. We also thank the Yakama Nation Tribal Council and Yakama Nation Fisheries for working with us to facilitate data collection and optimization of data usage according to their values and worldview.

54 ENVIRONMENTAL SCIENCES↗

Quantum-centric supercomputing for materials science: A perspective on challenges and future directions

Computational models are an essential tool for the design, characterization, and discovery of novel materials. Computationally hard tasks in materials science stretch the limits of existing high-performance supercomputing centers, consuming much of their resources for simulation, analysis, and data processing. Quantum computing, on the other hand, is an emerging technology with the potential to accelerate many of the computational tasks needed for materials science. In order to do that, the quantum technology must interact with conventional high-performance computing in several ways: approximate results validation, identification of hard problems, and synergies in quantum-centric supercomputing. Here in this paper, we provide a perspective on how quantum-centric supercomputing can help address critical computational problems in materials science, the challenges to face in order to solve representative use cases, and new suggested directions.

36 MATERIALS SCIENCE↗

2023 Project Peer Review Report

The Bioenergy Technologies Office (BETO) within the U.S. Department of Energy’s Office of Energy Efficiency and Renewable Energy supports the research, development, and demonstration (RD&D) of technologies aimed at mobilizing domestic renewable carbon resources for the reduction of greenhouse gas emissions across the U.S. economy. BETO systematically prioritizes RD&D into technology opportunities across a range of emerging scientific breakthroughs and technology readiness levels in the subprogram areas illustrated in Figure 1. This approach supports a diverse portfolio while developing the most promising and widely applicable technologies, testing technologies as integrated processes, and demonstrating integrated processes to support scale-up. These technologies will use a broad variety of renewable carbon resources to produce increasing volumes of biofuels and bioproducts. More information on BETO’s mission, goals, and strategic approaches can be found in the Bioenergy Technologies Office Multi-Year Program Plan. The biennial Peer Review process enables external stakeholders to provide feedback on the responsible use of taxpayer funding and develop recommendations for the most efficient and effective ways to accelerate the development of a bioenergy industry. This report includes the results of the Project Peer Review meeting held on April 3–7, 2023, in Denver, Colorado.

09 BIOMASS FUELS↗

ARM FY2026 Radar Plan

The U.S. Department of Energy (DOE) Atmospheric Radiation Measurement (ARM) User Facility maintains a suite of advanced atmospheric radar systems that serve as critical tools in ARM’s mission to provide continuous, high-quality observations for advancing the understanding and modeling of atmospheric processes. These radar systems enable detailed characterization of clouds, precipitation, and dynamic structures in the atmosphere, supporting a broad range of scientific applications. The number of deployed systems exceeds what current staffing levels can fully support for continuous 24/7/365 operation. As such, it is essential to have a clearly defined and community-informed plan that prioritizes radar operations and communicates ARM’s strategy for sustaining and evolving these observational assets. This FY2026 Radar Plan outlines ARM’s approach to managing its radar portfolio—balancing scientific impact, operational feasibility, and long-term sustainability. It reflects ARM’s continued commitment to delivering calibrated, well-documented radar data products that enable process-level studies and support the development and evaluation of weather and climate models. Through this plan, ARM aims to ensure transparency in decision-making, alignment with user needs, and support for innovative science across the facility’s fixed and mobile observatories. Given uncertainties around the Fiscal Year (FY) 2026 budget, this plan was developed to assume business as usual and will be updated as budgets and plans may change. It should be noted that, given the limited timeframe involved, this plan will be more succinct than previous plans.

47 OTHER INSTRUMENTATION↗

Validation of a Global Geospace Model With a Systems Science Approach Based on Canonical Correlation Analysis

A systems science approach based on canonical correlation analysis (CCA) is applied as a new, behavioral way to validate global geospace models. The biggest novelty of the technique is that it validates models at a system level, whereby a side‐by‐side comparison is performed of CCA applied to a 30‐day observational and the corresponding simulation data sets comprising quiet, moderate and active times. The simulation used the Multiscale Atmosphere‐Geospace Environment (MAGE) model. It is shown that (a) CCA must be combined with sensitivity analysis to be effective, (b) the MAGE model generally reproduces the observed behavior (more so for quieter time intervals), quantified by the intercorrelations between different variables and (c) the technique identifies the SuperMAG SML index as a quantity for which refinements of the model are needed.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Refining Planetary Boundary Layer Height Retrievals From Micropulse‐Lidar at Multiple ARM Sites Around the World

Abstract Knowledge of the planetary boundary layer height (PBLH) is crucial for various applications in atmospheric and environmental sciences. Lidar measurements are frequently used to monitor the evolution of the PBLH, providing more frequent observations than traditional radiosonde‐based methods. However, lidar‐derived PBLH estimates have substantial uncertainties, contingent upon the retrieval algorithm used. In addressing this, we applied the Different Thermo‐Dynamic Stabilities (DTDS) algorithm to establish a PBLH data set at five separate Department of Energy's Atmospheric Radiation Measurement sites across the globe. Both the PBLH methodology and the products are subject to rigorous assessments in terms of their uncertainties and constraints, juxtaposing them with other products. The DTDS‐derived product consistently aligns with radiosonde PBLH estimates, with correlation coefficients exceeding 0.77 across all sites. This study delves into a detailed examination of the strengths and limitations of PBLH data sets with respect to both radiosonde‐derived and other lidar‐based estimates of the PBLH by exploring their respective errors and uncertainties. It is found that varying techniques and definitions can lead to diverse PBLH retrievals due to the inherent intricacy and variability of the boundary layer. Our DTDS‐derived PBLH data set outperforms existing products derived from ceilometer data, offering a more precise representation of the PBLH. This extensive data set paves the way for advanced studies and an improved understanding of boundary‐layer dynamics, with valuable applications in weather forecasting, climate modeling, and environmental studies.

54 ENVIRONMENTAL SCIENCES↗

Characterizing Wet Season Precipitation in the Central Amazon Using a Mesoscale Convective System Tracking Algorithm

To comprehensively characterize convective precipitation in the central Amazon region, we utilize the Python FLEXible object TRacKeR (PyFLEXTRKR) to track mesoscale convective systems (MCSs) observed through satellite measurements and simulated by the Weather Research and Forecasting model at a convection-permitting resolution. This study spans a 2-month period during the wet seasons of 2014 and 2015. We observe a strong correlation between the MCS track density and accumulated precipitation in the Amazon basin. Key factors contributing to precipitation, such as MCS properties (number, size, rainfall intensity, and movement), are thoroughly examined. Our analysis reveals that while the overall model produces fewer MCSs with smaller mean sizes compared to observations, it tends to overpredict total precipitation due to excessive rainfall intensity for heavy rainfall events (≥10 mm hr –1 ). These biases in simulated MCS properties could vary with the constraints on the convective background environment. Moreover, while the wet bias from heavy (convective) rainfall outweighs the dry bias in light (stratiform) rainfall, the latter can be crucial, particularly when MCS cloud cover is significantly underestimated. A case study for 1 April 2014 highlights the influence of environmental conditions on the MCS lifecycle and identifies an unrealistic model representation in both stratiform and convective precipitation features.

54 ENVIRONMENTAL SCIENCES↗

HPC-Enabled Optimization of High Temperature Heat Exchangers (CRADA Final Report)

This project was a collaborative effort between Lawrence Livermore National Security, LLC (LLNS) as manager and operator of Lawrence Livermore National Laboratory (LLNL) and Materials Sciences, LLC, to develop a technology for design and optimization of heat exchangers using powerful desktop and laptop computers. The project was originally designated as a 12-month project, and consisted of 3# major tasks and the following 8# major deliverables: 1) CFD models of 3D heat exchangers based on existing and new geometry. 2) Validation against experimental data provided by MSC and published in the literature. 3) CFD models of 3D unit cells based on TPMS. 4) Surrogate models capable of delivering the gradients of the homogenized properties with respect to the parametrization. 5) 3D design methodology using TO algorithms. 6) Conventional reference and topology optimized designs. 7) 3D optimized designs stored in a 3D printer build format. 8) Verification of the improved performance. All of the deliverables for this project were successfully completed with two no-cost time extensions.

13 HYDRO ENERGY↗

Learning Constitutive Relations From Soil Moisture Data via Physically Constrained Neural Networks

Abstract The constitutive relations of the Richardson‐Richards equation encode the macroscopic properties of soil water retention and conductivity. These soil hydraulic functions are commonly represented by models with a handful of parameters. The limited degrees of freedom of such soil hydraulic models constrain our ability to extract soil hydraulic properties from soil moisture data via inverse modeling. We present a new free‐form approach to learning the constitutive relations using physically constrained neural networks. We implemented the inverse modeling framework in a differentiable modeling framework, JAX, to ensure scalability and extensibility. For efficient gradient computations, we implemented implicit differentiation through a nonlinear solver for the Richardson‐Richards equation. We tested the framework against synthetic noisy data and demonstrated its robustness against varying magnitudes of noise and degrees of freedom of the neural networks. We applied the framework to soil moisture data from an upward infiltration experiment and demonstrated that the neural network‐based approach was better fitted to the experimental data than a parametric model and that the framework can learn the constitutive relations.

54 ENVIRONMENTAL SCIENCES↗

How robust are estimates of key parameters in standard viral dynamic models?

Mathematical models of viral infection have been developed, fitted to data, and provide insight into disease pathogenesis for multiple agents that cause chronic infection, including HIV, hepatitis C, and B virus. However, for agents that cause acute infections or during the acute stage of agents that cause chronic infections, viral load data are often collected after symptoms develop, usually around or after the peak viral load. Consequently, we frequently lack data in the initial phase of viral growth, i.e., when pre-symptomatic transmission events occur. Missing data may make estimating the time of infection, the infectious period, and parameters in viral dynamic models, such as the cell infection rate, difficult. However, having extra information, such as the average time to peak viral load, may improve the robustness of the estimation. Here, we evaluated the robustness of estimates of key model parameters when viral load data prior to the viral load peak is missing, when we know the values of some parameters and/or the time from infection to peak viral load. Although estimates of the time of infection are sensitive to the quality and amount of available data, particularly pre-peak, other parameters important in understanding disease pathogenesis, such as the loss rate of infected cells, are less sensitive. Viral infectivity and the viral production rate are key parameters affecting the robustness of data fits. Fixing their values to literature values can help estimate the remaining model parameters when pre-peak data is missing or limited. We find a lack of data in the pre-peak growth phase underestimates the time to peak viral load by several days, leading to a shorter predicted growth phase. On the other hand, knowing the time of infection (e.g., from epidemiological data) and fixing it results in good estimates of dynamical parameters even in the absence of early data. While we provide ways to approximate model parameters in the absence of early viral load data, our results also suggest that these data, when available, are needed to estimate model parameters more precisely.

59 BASIC BIOLOGICAL SCIENCES↗

Airborne LiDAR to Improve Canopy Fuels Mapping for Wildfire Modeling

Increasing conflict between wildfire and the built environment has increased the need for more up-to-date and finer resolution canopy fuels data to improve wildfire modeling and associated risk forecasts. The US Forest Service and US Department of the Interior’s LANDFIRE product, which provides 30-m resolution canopy fuels data for the entire US, is one of the most widely used sources of fuels data. However, the last complete mapping effort for LANDFIRE is based on 2016 conditions, and subsequent updates reflect disturbances 1-2 years behind the release year. Airborne systems equipped with Light Detection and Ranging (LiDAR) sensors can be deployed to actively sense canopy structure and estimate canopy fuels data (cover, height, base height, bulk density) at finer resolutions. Canopy base height (CBH) and canopy bulk density (CBD) are difficult to measure both in the field and in LiDAR point clouds. Still, they are important for accurately modeling crown fires, which are often intense and difficult to contain. Additionally, point cloud datasets are large, and calculations require efficient utilization of computational resources. To address these challenges, we are working on an approach that uses openly available National Ecological Observatory Network (NEON) airborne LiDAR data, with calculations processed in the R programming language and parallelized through the lidR package. CBH and CBD are often derived from tree height, diameter at breast height, and species-specific allometries using the Fire and Fuels Extension of the Forest Vegetation Simulator (FFE-FVS). We aim to test if airborne LiDAR can estimate CBH and CBD without the use of empirical equations. Reliable estimates of canopy fuels data directly from airborne LiDAR could streamline quick, fine-resolution updates for use in wildfire behavior models.

54 ENVIRONMENTAL SCIENCES↗

Multi-Scale Modeling of the Evolution of Structure and Properties in Materials for Nuclear Energy Applications [Slides]

Nuclear energy is an important component of an overall strategy to address climate change. Idaho National Laboratory (INL) is the U.S. Department of Energy’s primary facility for research and development in nuclear science and technology for energy generation, supporting the improvement and life extension of the existing reactor fleet and the development and licensing of new reactor designs. Computational modeling is an important component of these activities, particularly in the area of materials for nuclear applications, where experimental data can be very challenging and expensive to acquire, and where data is especially scarce for new reactor designs. INL has used multi-scale modeling – linking atomistic, mesoscale, and engineering scales – to improve the ability to predict the performance of materials for nuclear energy applications. In this talk, I will give an overview of the approach and tools used, and several examples of application, including performance of nuclear fuels, understanding radiation-driven formation of nanoscale void and gas bubble superlattices, and powder densification through electric field assisted sintering.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Neural network representations of multiphase Equations of State

Abstract Equations of State model relations between thermodynamic variables and are ubiquitous in scientific modelling, appearing in modern day applications ranging from Astrophysics to Climate Science. The three desired properties of a general Equation of State model are adherence to the Laws of Thermodynamics, incorporation of phase transitions, and multiscale accuracy. Analytic models that adhere to all three are hard to develop and cumbersome to work with, often resulting in sacrificing one of these elements for the sake of efficiency. In this work, two deep-learning methods are proposed that provably satisfy the first and second conditions on a large-enough region of thermodynamic variable space. The first is based on learning the generating function (thermodynamic potential) while the second is based on structure-preserving, symplectic neural networks, respectively allowing modifications near or on phase transition regions. They can be used either “from scratch” to learn a full Equation of State, or in conjunction with a pre-existing consistent model, functioning as a modification that better adheres to experimental data. We formulate the theory and provide several computational examples to justify both approaches, highlighting their advantages and shortcomings.

Science & Technology - Other Topics↗

Data and scripts associated with “When do Riverine Systems 'Feel the Burn'? Simulating How Burn Extent and Severity Modulate Hydrologic Controls on Biogeochemical Export” (v2)

This data package is associated with the publication “When do Riverine Systems 'Feel the Burn'? Simulating How Burn Extent and Severity Modulate Hydrologic Controls on Biogeochemical Export” published in Water Resources Research (Wampler et al. 2025; preprint: https://doi.org/10.22541/essoar.174438106.63564767/v1). This study used the Soil and Water Assessment Tool (SWAT), a processed based model to explore the impacts of area burned and burn severity on streamflow, nitrate, and dissolved organic carbon (DOC) in two test basins: a semi-arid, mixed land use basin and a humid, primarily forested basin. We developed 1800 wildfire scenarios that we ran in each basin: 20 different burn extents (5 to 100% by 5%), 3 different burn severities (low, moderate, and high), and 30 different post-fire precipitation scenarios. We also ran an additional 30 scenarios associated with no wildfire for the 30 post-fire precipitation scenarios. For each scenario we were interested in the change in runoff ratio (streamflow) and average concentration and annual loads (nitrate and DOC) across the wildfire scenarios. This data package contains the data and scripts required to build SWAT models for the two test basins, create and run the wildfire scenarios, and generate the data summaries and figures used in the associated manuscript. This data package was originally published in March 2025. It was updated in January 2026 (v2; new and modified files) to include the final files after the manuscript went through reviews. See the change history section below for more details. For details on how to navigate data packages generated by this project, see https://data.ess-dive.lbl.gov/portals/PNNLRiverCorridorSFA/About.

54 ENVIRONMENTAL SCIENCES↗