Search NASA⌕ Search

SEARCH · Search NASA

Results for “Common data models”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

Shift Happens: Building Robust AI Models with Domain Adaptation

Artificial Intelligence (AI) is revolutionizing physics research—from probing the large-scale structure of the Universe to modeling subatomic interactions and fundamental forces. Yet, a major challenge persists: AI models trained on simulations or old experiment / astronomical survey often perform poorly when applied to new data—exposing issues of dataset (domain) shift, model robustness, and uncertainty in predictions. This summer school session will introduce students to common challenges in applying AI across domains and present solutions based on domain adaptation—a set of techniques designed to improve model generalization under domain shift. We will cover foundational ideas, practical strategies, and current research frontiers in this area. Through examples in astrophysics, we'll explore how domain adaptation can help bridge the gap between synthetic and real-world data, improve trust in model outputs, and advance scientific discovery. The concepts discussed are broadly applicable across physics and other scientific disciplines, making this a valuable topic for anyone interested in building robust, transferable AI models for science.

Ciprijanovic, A. [Fermilab] (ORCID:000000031281719↗

First Double-Differential Cross Section Measurement of Neutral-Current 𝜋 0 Production in Neutrino-Argon Scattering in the MicroBooNE Detector

We report the first double-differential cross section measurement of neutral-current neutral pion (NC⁢𝜋 0 ) production in neutrino-argon scattering, as well as single-differential measurements of the same channel in terms of final states with and without protons. The kinematic variables of interest for these measurements are the 𝜋 0 momentum and the 𝜋 0 scattering angle with respect to the neutrino beam. A total of 4971 candidate NC⁢𝜋 0 events fully contained within the MicroBooNE detector are selected using data collected at a mean neutrino energy of ∼0.8 GeV from 6.4 ×10 20 protons on target from the Booster Neutrino Beam at the Fermi National Accelerator Laboratory. After extensive data-driven model validation to ensure unbiased unfolding, the Wiener-singular-value-decomposition method is used to extract nominal flux-averaged cross sections. The results are compared to predictions from commonly used neutrino event generators, which tend to overpredict the measured NC⁢𝜋 0 cross section, especially in the 0.2–0.5 GeV/c 𝜋 0 momentum range and at forward scattering angles. Events with at least one proton present in the final state are also underestimated. This data will help improve the modeling of NC⁢𝜋 0 production, which represents a major background in measurements of charge-parity violation in the neutrino sector and in searches for new physics beyond the standard model.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Alabama Carbon Storage: Bringing Data to the People

The Gulf Coastal Plain of Alabama has proven potential for geologic carbon storage and current interest in the area for large carbon capture and storage (CCS) projects is high. Extensive CCS relevant data exist in the records of the Geological Survey of Alabama and State Oil and Gas Board of Alabama, however, most of this data is not publicly available or is scattered in separate databases, file cabinets, and tables in publications. The “Alabama Carbon Storage: Data Sharing and Engagement” (ACS-DSE) project seeks to accelerate the responsible development of large CCS projects in the Gulf Coastal Plain of Alabama and offshore in state waters through a publicly accessible database of geologic carbon storage models and data across the region. The ACS-DSE draws on the over 150 years of geologic research and over 20 years of experience in CCS research to place relevant geologic, geophysical, and infrastructure data on a single web platform. Datasets available will include formation depths and elevations, geologic structures, reservoir properties, digital well logs (LAS files), existing penetrations, and geologic models. In addition to downloadable datasets, links to CCS related regulatory agencies and other sources of information will be included (for example, Class VI UIC permitting regulations and pipeline regulations). By making these datasets and models available in commonly used formats on a public website, the project will increase transparency in decision making and decrease the data acquisition time for industry.

01 COAL, LIGNITE, AND PEAT↗

Radiative impact of record-breaking wildfires from integrated ground-based data

The radiative effects of wildfires have been traditionally estimated by models using radiative transfer calculations. Assessment of model-predicted radiative effects commonly involves information on observation-based aerosol optical properties. However, lack or incompleteness of this information for dense plumes generated by intense wildfires reduces substantially the applicability of this assessment. Here we introduce a novel method that provides additional observational constraints for such assessments using widely available ground-based measurements of shortwave and spectrally resolved irradiances and aerosol optical depth (AOD) in the visible and near-infrared spectral ranges. We apply our method to quantify the radiative impact of the record-breaking wildfires that occurred in the Western US in September 2020. For our quantification we use integrated ground-based data collected at the Atmospheric Measurements Laboratory in Richland, Washington, USA with a location frequently downwind of wildfires in the Western US. We demonstrate that remarkably dense plumes generated by these wildfires strongly reduced the solar surface irradiance (up to 70% or 450 Wm -2 for total shortwave flux) and almost completely masked the sun from view due to extremely large AOD (above 10 at 500 nm wavelength). We also demonstrate that the plume-induced radiative impact is comparable in magnitude with those produced by a violent volcano eruption occurred in the Western US in 1980 and continental cumuli.

54 ENVIRONMENTAL SCIENCES↗

Joint Modeling of Wind Speed and Wind Direction Through a Conditional Approach

Atmospheric near surface wind speed and wind direction play an important role in many applications, ranging from air quality modeling, building design, wind turbine placement to climate change research. It is therefore crucial to accurately estimate the joint probability distribution of wind speed and direction. In this work, we develop a conditional approach to model these two variables, where the joint distribution is decomposed into the product of the marginal distribution of wind direction and the conditional distribution of wind speed given wind direction. To accommodate the circular nature of wind direction, a von Mises mixture model is used; the conditional wind speed distribution is modeled as a directional dependent Weibull distribution via a two-stage estimation procedure, consisting of a directional binned Weibull parameter estimation, followed by a harmonic regression to estimate the dependence of the Weibull parameters on wind direction. A Monte Carlo simulation study indicates that our method outperforms two other approaches in estimation efficiency: one that utilizes periodic spline quantile regression and another that generates data from the commonly used Abe-Ley distribution for cylindrical data. We illustrate our method by using the output from a regional climate model to investigate how the joint distribution of wind speed and direction may change under some future climate scenarios. Our method indicates significant changes in the variation of wind speed with respect to some directions.

17 WIND ENERGY↗

Opportunities for Earth Observation to Inform Risk Management for Ocean Tipping Points

Abstract As climate change continues, the likelihood of passing critical thresholds or tipping points increases. Hence, there is a need to advance the science for detecting such thresholds. In this paper, we assess the needs and opportunities for Earth Observation (EO, here understood to refer to satellite observations) to inform society in responding to the risks associated with ten potential large-scale ocean tipping elements: Atlantic Meridional Overturning Circulation; Atlantic Subpolar Gyre; Beaufort Gyre; Arctic halocline; Kuroshio Large Meander; deoxygenation; phytoplankton; zooplankton; higher level ecosystems (including fisheries); and marine biodiversity. We review current scientific understanding and identify specific EO and related modelling needs for each of these tipping elements. We draw out some generic points that apply across several of the elements. These common points include the importance of maintaining long-term, consistent time series; the need to combine EO data consistently with in situ data types (including subsurface), for example through data assimilation; and the need to reduce or work with current mismatches in resolution (in both directions) between climate models and EO datasets. Our analysis shows that developing EO, modelling and prediction systems together, with understanding of the strengths and limitations of each, provides many promising paths towards monitoring and early warning systems for tipping, and towards the development of the next generation of climate models.

Wood, Richard A. (ORCID:0000000239609513)↗

Comparison of removal and spatial mark‐resight models for estimating wild pig density

Density estimation is critical to effectively manage invasive species and elucidate areas of highest concern. For wild pigs (Sus scrofa), the ability to estimate density is complicated because of their variable home range sizes and social structure. Common methods for estimating density (e.g., mark-recapture) may be unsuitable in management applications because additional data needs to be collected before and after management. Removal models offer a suitable alternative to estimate density changes following management and can be applied broadly across areas where management of wild pigs is ongoing. We collected wild pig removal and camera trap data from 25 private properties ranging in size from approximately 0.5 km 2 to 95 km 2 across 3 ecoregions in South Carolina, USA, from 2020–2023. We compared factors affecting consistency and precision of property-level density estimates between removal and spatial mark-resight (SMR) models. In general, excluding 1 large outlier, density estimates from removal models were between 0.60 and 15.85 wild pigs/km 2 (median = 5.34) with a median coefficient of variation (CV) of 0.76 and 95% confidence intervals for the CV between 0.70 and 0.94. Similarly, excluding 1 large outlier, density estimates from SMR were between 0.22 and 30.97 wild pigs/km 2 (median = 5.48) with a median CV of 0.39 and 95% confidence intervals for the CV between 0.38 and 1.20. We found the precision of removal models was affected primarily by the number of wild pigs dispatched in the removal period (3 months) and the ecoregion in which they were removed. None of the covariates, including the number of recaptures (a corresponding measure of sample size), influenced precision of the SMR models, although recaptures did influence the density estimates. At the individual property level, density estimates from our 2 estimators were dissimilar from each other in approximately 80% of instances, although none of the covariates we examined influenced dissimilarity. Our results provide unique insight into how sample size affects density estimates using 2 common methods and into novel SMR models that incorporate both marked and unmarked detections. In addition, the density estimates in this study can be used as a reference for wild pig densities in common land cover types throughout the southeastern United States.

60 APPLIED LIFE SCIENCES↗

Model Data Archive for Manuscript Titled "Evaluation of a Coupled Surface–Subsurface Hydrologic Model Using Dense Water‑Level Sensors in a Mixed Urban–Rural Watershed"

This archive provides scripts, input files, and datasets used for the implementation and evaluation of a fully coupled surface–subsurface hydrologic model in the Neches River Basin, southeast Texas. The study uses the Advanced Terrestrial Simulator (ATS) to simulate coupled surface–subsurface hydrologic processes over a mixed urban–rural watershed and evaluates model performance using a dense network of 136 in situ water-level sensors, nine U.S. Geological Survey (USGS) stream gauges, and SSEBop-derived evapotranspiration estimates during the period October 2014–June 2024. The workflow is implemented primarily in Python 3 using the Watershed Workflow package. The Jupyter notebooks can be executed using open-source software such as Anaconda JupyterLab or Visual Studio Code. Other data files include TXT, CSV, XML, SHP, TIF, NetCDF, HDF5, and ExodusII files, which can be processed using the provided Python scripts. ATS input files are provided in XML format and can be edited using any commonly used text editor. This archive contains: *Scripts and input files used to generate the ATS model setup, including watershed discretization, mesh generation, parameter mapping, and model configuration. *Jupyter notebooks used for preprocessing observational data, evaluating streamflow, water levels, and evapotranspiration, computing performance metrics, and generating the figures presented in the manuscript. *ATS simulation outputs and processed observational datasets, including OneRain and DD6 water-level sensors, USGS streamflow observations, GIS data, and supporting spatial datasets used throughout the study.

Dense water-level sensor network↗

A flexible class of priors for orthonormal matrices with basis function-specific structure

Statistical modeling of high-dimensional matrix-valued data motivates the use of a low-rank representation that simultaneously summarizes key characteristics of the data and enables dimension reduction. Low-rank representations commonly factor the original data into the product of orthonormal basis functions and weights, where each basis function represents an independent feature of the data. However, the basis functions in these factorizations are typically computed using algorithmic methods that cannot quantify uncertainty or account for basis function correlation structure a priori. While there exist Bayesian methods that allow for a common correlation structure across basis functions, empirical examples motivate the need for basis function-specific dependence structure. We propose a prior distribution for orthonormal matrices that can explicitly model basis function-specific structure. The prior is used within a general probabilistic model for singular value decomposition to conduct posterior inference on the basis functions while accounting for measurement error and fixed effects. We discuss how the prior specification can be used for various scenarios and demonstrate favorable model properties through synthetic data examples. Finally, we apply our method to two-meter air temperature data from the Pacific Northwest, enhancing our understanding of the Earth system’s internal variability.

97 MATHEMATICS AND COMPUTING↗

Modeling Distributed Generation in California

In support of analysis for the biennial Integrated Energy Policy Report, the California Energy Commission and the National Renewable Energy Laboratory have partnered to study the growth of distributed energy resources in California. This study involves the use of National Renewable Energy Laboratory's Distributed Generation Market Demand model, available at https://www.nrel.gov/analysis/dgen/, to project statewide adoption of distributed photovoltaics and paired storage. Key outcomes of the collaboration include: • Improved representation of California building stock, load profiles, historical adoption, and tariffs, including the net billing tariff, in the dGen model; • Trained CEC staff members to use and adapt the dGen model for their specific needs; • Developed a methodology for representing emerging consumer segments to potentially adopt distributed energy resources, including low-income, multifamily, and renter-occupied buildings; • Forecasted solar photovoltaic and paired storage growth in California using a common set of modeling parameters. This report describes the multiyear effort, which includes a discussion of: • Methodology and data employed in adapting the Distributed Generation Market Demand model for California to forecast solar photovoltaic and storage statewide through 2040; • Steps taken to modify the base model to forecast solar photovoltaic adoption in emerging market segments such as multifamily or renter-occupied homes or both; • Future enhancements of the model.

14 SOLAR ENERGY↗

Extraction and Analysis of Time Series Data from Building Automation Systems Using Large Language Models

Semantic schemas like Haystack 4, Brick and ASHRAE standard 223 enable the structured, standardized, and machine-readable representation of building data, facilitating interoperability, data integration, and advanced analytics. However, extracting information from these models requires specialized expertise in SPARQL and other programming languages, skills that are not commonly found among building professionals. Recent advancements in Large Language Models (LLMs), such as ChatGPT, enable the construction of queries using natural language, making it easier for individuals to interact with these systems in a manner that resembles everyday speech. However, these methods have not yet been tested on building semantic ontologies. This paper introduces a novel workflow and tool for enabling users to ask questions about a specific building's data, using natural language and receive answers automatically generated by GPT-4o. Our approach integrates semantic ontologies with advanced LLM capabilities to automate three critical steps: (1) generating SPARQL queries to retrieve time series references from ontological models, (2) extracting the corresponding time series data from the Building Automation System, and (3) performing computations and visualizations tailored to the user's query. The proposed method simplifies access to BAS data, allowing both domain experts and non-specialists to conduct sophisticated analyses without needing extensive technical knowledge of semantic web technologies. By demonstrating this pipeline, we facilitate more accessible and scalable data-driven decision-making in building operations and management.

Mulayim, Ozan Baris↗

Extraction and Analysis of Time Series Data from Building Automation Systems Using Large Language Models

Semantic schemas like Haystack 4, Brick and ASHRAE standard 223 enable the structured, standardized, and machine-readable representation of building data, facilitating interoperability, data integration, and advanced analytics. However, extracting information from these models requires specialized expertise in SPARQL and other programming languages, skills that are not commonly found among building professionals. Recent advancements in Large Language Models (LLMs), such as ChatGPT, enable the construction of queries using natural language, making it easier for individuals to interact with these systems in a manner that resembles everyday speech. However, these methods have not yet been tested on building semantic ontologies. This paper introduces a novel workflow and tool for enabling users to ask questions about a specific building's data, using natural language and receive answers automatically generated by GPT-4o. Our approach integrates semantic ontologies with advanced LLM capabilities to automate three critical steps: (1) generating SPARQL queries to retrieve time series references from ontological models, (2) extracting the corresponding time series data from the Building Automation System, and (3) performing computations and visualizations tailored to the user's query. The proposed method simplifies access to BAS data, allowing both domain experts and non-specialists to conduct sophisticated analyses without needing extensive technical knowledge of semantic web technologies. By demonstrating this pipeline, we facilitate more accessible and scalable data-driven decision-making in building operations and management.

Mulayim, Ozan Baris↗

Robust Solution Verification Experiments on Nonuniform Meshes

The activities of verification, validation, and uncertainty quantification (VVUQ) provide a comprehensive means to assess the credibility of computational models. Within VVUQ, solution verification assesses numerical errors and evaluates whether the simulation is sufficiently accurate for its intended applications. As computational modeling gains traction in the development of complex, high-consequence systems, the need for robust solution verification intensifies, particularly because experimental data for these systems are often limited. This work examines improvements in the robustness of Richardson extrapolation (RE), a method commonly used in solution verification to study the discretization error of computational models using a power law. Nonuniform mesh refinement is discussed alongside other pollutants that affect the robustness of the power law model. Maximum likelihood estimation (MLE) is proposed as a robust strategy to address the uncertainty generated by nonuniform mesh refinement. An exploratory computational fluid dynamics (CFD) study of a 2D planar Poiseuille flow is conducted to determine if nonuniform mesh noise can be modeled with this MLE approach for more robust RE.

Weinmeister, Justin [ORNL] (ORCID:0000000160090237↗

Explosion Detection Using Smartphones: Ensemble Learning with the Smartphone High-Explosive Audio Recordings Dataset and the ESC-50 Dataset

Explosion monitoring is performed by infrasound and seismoacoustic sensor networks that are distributed globally, regionally, and locally. However, these networks are unevenly and sparsely distributed, especially at the local scale, as maintaining and deploying networks is costly. With increasing interest in smaller-yield explosions, the need for more dense networks has increased. To address this issue, we propose using smartphone sensors for explosion detection as they are cost-effective and easy to deploy. Although there are studies using smartphone sensors for explosion detection, the field is still in its infancy and new technologies need to be developed. We applied a machine learning model for explosion detection using smartphone microphones. The data used were from the Smartphone High-explosive Audio Recordings Dataset (SHAReD), a collection of 326 waveforms from 70 high-explosive (HE) events recorded on smartphones, and the ESC-50 dataset, a benchmarking dataset commonly used for environmental sound classification. Two machine learning models were trained and combined into an ensemble model for explosion detection. The resulting ensemble model classified audio signals as either “explosion”, “ambient”, or “other” with true positive rates (recall) greater than 96% for all three categories.

45 MILITARY TECHNOLOGY, WEAPONRY, AND NATIONAL DEF↗

Monsoonal MCS Initiation, Rainfall, and Diurnal Gravity Waves over the Bay of Bengal: Observation and a Linear Model

Abstract Previous observational studies have indicated that mesoscale convective systems (MCSs) contribute the majority of precipitation over the Bay of Bengal (BoB) during the summer monsoon season, yet their initiation and propagation remain incompletely understood. To fill this knowledge gap, we conducted a comprehensive study using a combination of 20-yr satellite observations, MCS tracking, reanalysis data, and a theoretical linear model. Satellite observations reveal clear diurnal propagation signals of MCS initiation frequency and rainfall from the west coast of the BoB toward the central BoB, with the MCS rainfall propagating slightly slower than the MCS initiation frequency. Global reanalysis data indicate a strong association between the offshore-propagating MCS initiation frequency/rainfall and diurnal low-level wind perturbations, implying the potential role of gravity waves. To verify the hypothesis, we developed a 2D linear model that can be driven by realistic meteorological fields from reanalysis. The linear model realistically reproduces the characteristics of offshore-propagating diurnal wind perturbations. The wind perturbations, as well as the offshore propagation signals of MCS initiation frequency and rainfall, are associated with diurnal gravity waves emitted from the coastal regions, which in turn are caused by the diurnal land–sea thermal contrast. The ambient wind speed and vertical wind shear play crucial roles in modulating the timing, propagation, and amplitude of diurnal gravity waves. Using the linear model and satellite observations, we further show that the stronger monsoonal flows lead to faster offshore propagation of diurnal gravity waves, which subsequently control the offshore propagation signals of MCS initiation and rainfall. Significance Statement Rainfall over the Bay of Bengal (BoB) is primarily contributed by large and organized rainfall systems in the summer monsoon season. During this season, these systems are commonly observed over the east coast of India around midnight, the western BoB in the morning, and the central BoB in the afternoon. This eastward rainfall propagation is confirmed by observations, reanalysis data, and a theoretical model to have a strong association with atmospheric diurnal gravity waves. These waves are caused by land–sea thermal contrast and can trigger rainfall systems over the offshore regions. We also found that the diurnal gravity waves, as well as their triggered rainfall systems, can be greatly modulated by the large-scale monsoonal flows. All the above findings improve our understanding of diurnal rainfall cycle over the BoB and may contribute to the future improvement of rainfall forecast over the region.

Meteorology & Atmospheric Sciences↗

Development of a 95-Year Solar Dataset for Resource Adequacy Studies

Long-term high-resolution solar data provides enhanced understanding of variability of solar generation and enhances our ability to develop strategies for a resilient and reliable electric grid under high deployment of solar energy. Therefore, it is important to develop long-term synthetic datasets that can provide multiple occurrences of various severe weather scenarios that are expected to test the limits of resource adequacy under scenarios contain various energy generation sources. Examples of such scenarios could be long periods of high temperatures when demand for electricity is high or periods where high winds could lead to a shut-down of transmission lines for long periods of time to ensure fire safety. NREL has developed the first version of such a dataset covering a 95-year period covering 2006-2100 at a 4km hourly resolution. This dataset contains all variables necessary to calculate solar generation. During development of this dataset, we focused on creating unbiased, high-resolution solar irradiance through statistical downscaling methods, using Regional Climate Model (RCM) simulations from the North American Coordinated Regional Climate Downscaling Experiment (NA-CORDEX) as input. The National Solar Radiation Database (NSRDB) containing over 25 years of observations was used to calibrate the statistical downscaling models. This presentation will outline the primary steps in developing this dataset, including (1) regridding RCM data to a common grid at 20-km resolution, (2) correcting RCM biases with NSRDB, (3) applying temporal and spatial downscaling methods to generate high-resolution (4-km, hourly) solar and ancillary data. Additionally, we will present an evaluation of the downscaled data against the NSRDB across various zones in the CONUS. Lastly, we will present a user guide for accessing the datasets.

14 SOLAR ENERGY↗

Moisture Metaphenome Incubation Analysis Results

The Birch effect, a pulse of CO2 release that occurs when dry soil is rewet, is commonly observed, yet the underlying biogeochemistry remains elusive. Using multi-omics data, real-time mass spectrometry and modeling approaches, we investigated the molecular response to rewetting of a soil microbiome exposed to drought for one and two weeks. The microbiome response was evaluated through analysis of transcript, protein, metabolite, and respiration profiles and metabolic modeling using an enhanced version of the Metabolite Expression Metabolic Network Integration for Pathway Identification and Selection (MEMPIS) algorithm (Roy Chowdhury et al, mSystems, 2019).

Lipton, Mary S [Pacific Northwest National Laborat↗

Layers Can Be Deceiving: A Hopping Model for Small Molecule Diffusion in TATB Crystal

Sorption of small molecule gases in materials can play a significant role in their long‐term stability and compatibility within multi‐material assemblies. While many material properties of the insensitive high explosive TATB (1,3,5‐triamino‐2,4,6‐trinitrobenzene) are well understood, very little is known regarding its permeability to gases. TATB crystal exhibits a graphitic‐like layered packing structure that evokes a mental schema in which the layers form nanoscopic channels, but it is unclear whether this structure promotes gas transport. Here, we use molecular dynamics (MD) simulations to predict transport of small molecules through TATB single crystal. An approach to fit classical force fields is developed to model TATB interactions with H 2 O, He, Ne, and Ar, which is then combined with steered MD to probe gas transport along selected directions in the crystal. We find that small molecule transport occurs via a hopping mechanism that exhibits distinct jumps between interstitial sites and is substantially faster normal to the layers as compared to through them. This result stems from the finding that intralayer junctions between adjacent TATB molecules are the most stable interstitial sites and that energetic barriers are lower for hopping between adjacent layers. An empirical model for diffusion rate based on the MD data shows that the rate decays exponentially with increasing molecular radius and is negligibly small for all molecules larger than He, including common atmospheric gases. These findings have implications for the interpretation of experiments that measure surface area, material response to extreme conditions, and are expected to help constrain models for material aging.

36 MATERIALS SCIENCE↗