Search NASA⌕ Search

SEARCH · Search NASA

Results for “Common data models”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 253 records · Page 14

Assessment of Bird Strike Likelihood to Refine Bird Strike Risk Models

In its most basic form, bird strike risk is comprised of a frequency component that reflects the likelihood of a collision and a severity component that reflects the cost (monetary or otherwise) of the incident. The bird strike risk model currently used by United State Department of Agriculture (USDA) Wildlife Services to evaluate the risk posed by individual bird species at airports and establish priorities for management was developed in 2018. The model uses airport-specific data on the number of reported strikes for a species recorded in the Federal Aviation Administration (FAA)’s National Wildlife Strike Database as a measure of frequency and the species’ relative hazard score as a measure of severity. The model was tested against independent data, found to perform well overall, and is being implemented widely across the United States. However, the model has limitations, including that species known to pose risk to aircraft locally, but not present in the strike record database, are not reflected as a major component of risk. Standard bird survey methodology commonly used at airports (e.g. point counts or transects) potentially can be used to complement wildlife strike records to calculate frequency or relative abundance of species. However, these methods generally focus on airport-wide population estimation and often ignore vital information that contributes to the true likelihood of a strike, such as use of runway protection zones and other critical areas, and spatial and temporal overlap with departing or approaching aircraft. As such, a more detailed understanding of space use by birds across landcovers and population fluctuations across the year is needed to accurately estimate the likelihood of bird strikes at airports. In this manuscript, we will review the extant risk model, including a discussion on its limitations. We then discuss approaches for refining our understanding of strike likelihood and briefly touch on needs for estimating probability of strike severity (cost).

bird strike, aircraft collision, damage by wildlif↗

Common femtoscopic hadron-emission source in pp collisions at the LHC

The femtoscopic study of pairs of identical pions is particularly suited to investigate the effective source function of particle emission, due to the resulting Bose–Einstein correlation signal. In small collision systems at the LHC, pp in particular, the majority of the pions are produced in resonance decays, which significantly affect the profile and size of the source. In this work, we explicitly model this effect in order to extract the primordial source in pp collisions at $\sqrt{s}$ = 13 TeV from charged π–π correlations measured by ALICE. We demonstrate that the assumption of a Gaussian primordial source is compatible with the data and that the effective source, resulting from modifications due to resonances, is approximately exponential, as found in previous measurements at the LHC. The universality of hadron emission in pp collisions is further investigated by applying the same methodology to characterize the primordial source of K–p pairs. The size of the primordial source is evaluated as a function of the transverse mass (m T ) of the pairs, leading to the observation of a common scaling for both π–π and K–p, suggesting a collective effect. Further, the present results are compatible with the m T scaling of the p–p and p–Λ primordial source measured by ALICE in high multiplicity pp collisions, providing additional evidence for the presence of a common emission source for all hadrons in small collision systems at the LHC. This will allow the determination of the source function for any hadron–hadron pairs with high precision, granting access to the properties of the possible final-state interaction among pairs of less abundantly produced hadrons, such as strange or charmed particles.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Using MARCUS, MICRE, and COMBLE data to improve understanding and modeling of cloud, aerosol, and boundary layer processes at high-latitudes

Because it is believed general circulation (GCM) and numerical weather prediction (NWP) models underestimate shortwave radiation over the Southern Ocean (SO) due to inadequate representations of boundary layer (BL) and cloud processes, it is critical to improve our understanding of key aerosol, cloud, precipitation, and BL processes. Over the north Atlantic Ocean (NA), cold-air outbreaks are common, yet few studies of the aerosol and environmental controls of the associated convective BL clouds exist as needed to develop and evaluate GCM representations. Although processes cannot be observed, cloud and aerosol properties can be measured in-situ or remotely retrieved, which when combined with numerical simulations enable process level understanding required to improve model representations.

54 ENVIRONMENTAL SCIENCES↗

The importance of accounting for the Tolman correction to surface tension for nucleation and growth modeling of Fe clusters

Gibbs free energies of clusters are required for predictive modeling of cluster growth during condensation of a cooling vapor. Here, we present a straightforward method of calculating free energies of cluster formation using the data from molecular dynamics (MD) simulations. We apply this method to iron clusters having from 2 to 100 atoms. The energies obtained are verified by comparing to an MD-simulated equilibrium cluster size distribution in a sub-saturated vapor. We show that these free energies differ significantly from those obtained with a commonly used spherical cluster approximation, which relies on a surface tension coefficient of a flat surface, as it is used in the classical nucleation theory (CNT). We show that the spherical cluster approximation in CNT can be improved by using a cluster-size-dependent Tolman correction for the surface tension. The Tolman length and effective surface tension values were derived for iron clusters, and they significantly differ from the commonly used experimentally measured values. This improved approximation does not account for geometric magic number effects responsible for spikes and troughs in densities of neighbor cluster sizes. Nonetheless, it allows to more accurately model cluster formation from a cooling vapor. It better reproduces the condensation timeline, overall shape of the cluster size distribution, average cluster size, and the distribution width. In contrast, using a constant surface tension coefficient (as done in CNT) resulted in incorrect condensation dynamics and cluster size distributions. The analytical expression for cluster nucleation rate from CNT was updated to account for the size-dependence of cluster surface tension.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

NuGraph2: A Graph Neural Network for Neutrino Event Reconstruction

Neutrino experiments are set to probe some of the most important open questions in physics, from CP violation and the nature of dark matter. The technology of choice for many of these experiments is the liquid argon time projection chamber (LArTPC). In current LArTPC experiments, reconstruction performance often represents a limiting factor for the sensitivity. New developments are therefore needed to unlock the full potential of LArTPC experiments. NuGraph2 is a state of the art Graph Neural Network for reconstruction of data in LArTPC experiments [https://arxiv.org/abs/2403.11872]. NuGraph2 utilizes a heterogeneous graph structure, with separate subgraphs of 2D nodes (hits in each plane) connected across planes via 3D nodes (space points). The model provides a consistent description of the neutrino interaction across all planes. NuGraph2 is a multi-purpose network, with a common message-passing attention engine connected to multiple decoders with different classification or regression tasks. These include the classification of detector hits according to the particle type that produced them (semantic segmentation) and the separation of hits from the neutrino interaction from hits due to noise or cosmic-ray background. Additional decoders are being developed, performing tasks such as the regression of the neutrino interaction vertex position. Performance results will be presented based on publicly available samples from MicroBooNE. These include both physics performance metrics, achieving 95% accuracy for semantic segmentation and 98% classification of neutrino hits, as well as computational metrics for training and for inference on CPU or GPU. The status of the NuGraph integration in the LArSoft software framework will be presented, as well as initial studies about model interpretability and injection of domain knowledge.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Spectroscopic Characterization of redMaPPer Galaxy Clusters with DESI

Optical galaxy cluster identification algorithms such as redMaPPer promise to enable an array of astrophysical and cosmological studies, but suffer from biases whereby galaxies in front of and behind a galaxy cluster are mistakenly associated with the primary cluster halo. These projection effects caused by irreducible photometric redshift uncertainty must be quantified to facilitate the use of optical cluster catalogues. We present measurements of galaxy cluster projection effects and velocity dispersion using spectroscopy from the Dark Energy Spectroscopic Instrument. Our findings are as follows: we confirm that the fraction of redMaPPer putative member galaxies mistakenly associated with cluster haloes is richness dependent, being more than twice as large at low richness than high richness; we present the first spectroscopic evidence of an increase in projection effects with increasing redshift, by as much as 25 per cent from $z\sim 0.1$ to $z\sim 0.2$; moreover, we find qualitative evidence for luminosity dependence in projection effects, with fainter galaxies being more commonly far behind clusters than their bright counterparts; finally, we fit the scaling relation between measured mean spectroscopic richness and velocity dispersion, finding an implied linear scaling between spectroscopic richness and halo mass. We discuss further directions for the application of spectroscopic data sets to improve use of optically selected clusters to test cosmological models.

clusters↗

Online learning of quadratic manifolds from streaming data for nonlinear dimensionality reduction and nonlinear model reduction

Here, this work introduces an online greedy method for constructing quadratic manifolds from streaming data, designed to enable in situ analysis of numerical simulation data on the Petabyte scale. Unlike traditional batch methods, which require all data to be available upfront and take multiple passes over the data, the proposed online greedy method incrementally updates quadratic manifolds in one pass as data points are received, eliminating the need for expensive disk input/output operations as well as storing and loading data points once they have been processed. A range of numerical examples demonstrate that the online greedy method learns accurate quadratic manifold embeddings while being capable of processing data that far exceed common disk input/output capabilities and volumes as well as main-memory sizes.

97 MATHEMATICS AND COMPUTING↗

Implementation of an extensible property modeling framework in ESPEI with applications to molar volume and elastic stiffness models

Property models are becoming more widely adopted by commercial Calphad databases, but they are not nearly as common in non-commercial or traditional academic Calphad databases. A primary driver is that user-friendly Calphad modeling tools that support property models are not widely available. Here we present new property modeling capabilities that have been implemented in ESPEI (the Extensible, Self-optimizing Phase Equilibrium Infrastructure). These capabilities include both generating property model parameters from data and improvements to the algorithmic selection of the most appropriate model from a series of candidates. Additionally, two illustrative examples are given that use ESPEI to fit different property models. First, we generate molar volume model parameters for Group IV, V, and VI refractory BCC alloys based on the model by Lu et al. (2005). Second, we demonstrate the extensibility of ESPEI’s property modeling capabilities by implementing a custom PyCalphad model for BCC elastic stiffness parameters to generate and compare parameters to the ones assessed by Marker et al. (2018) using the same data. Property models generated by ESPEI can be used in PyCalphad or further optimized with uncertainty quantification using ESPEI.

36 MATERIALS SCIENCE↗

Status of ICARUS NUMI Interaction Cross-Section Analysis

The ICARUS experiment, utilizing Liquid Argon Time Projection Chamber (LAr TPC) technology, has been installed at Fermilab in Chicago, Illinois, following its initial operation in Italy and subsequent refurbishment at CERN. ICARUS completed commissioning in June 2022. Currently, the experiment is in the phase of analyzing data from its two runs of physics data acquisition and gearing up for the third run. While its primary objective is to function as the far detector of the Short Baseline Neutrino program (SBN), seeking sterile neutrino signatures, ICARUS also offers diverse physics capabilities, including searches beyond the standard model and measurements of cross-sections. In addition to being exposed to the common Booster Neutrino (BNB) beamline, ICARUS also receives off-axis neutrinos from the Main Injector (NuMI) beam. Due to the off-axis angle between NuMI and ICARUS, coupled with contributions from both pion and kaon decays to neutrino fluxes, interactions of NuMI neutrinos within ICARUS can be detected over a range of several GeV in energy. These interactions present opportunities for crucial cross-section measurements and model tests within an energy range that overlaps both the SBN oscillation search and a portion of the DUNE spectrum. This poster presentation will delve into our efforts to conduct a muon-neutrino cross-section measurement, where the signal is defined by events with no pions produced in the final state of the interaction, along with some preliminary muon-neutrino inclusive measurements. Additionally, it will provide updates on the current status and future plans, including reconstruction, selection, and analysis procedures.

43 PARTICLE ACCELERATORS↗

Development of an Unbiased Future Solar Dataset for Solar Resource Adequacy Research Over CONUS

A high-resolution, long-term solar dataset is essential for capturing the variability of solar energy resources and informing strategies to ensure grid reliability and resilience in systems with high levels of solar energy integration. This study focuses on generating unbiased, high-resolution projections of solar irradiance through a statistical downscaling framework, using Earth system model (ESM) simulations obtained from the North American Coordinated Regional Climate Downscaling Experiment (NA-CORDEX). The National Solar Radiation Database (NSRDB) is used to calibrate statistical downscaling models. The newly developed dataset provides solar irradiance, surface air temperature, and surface wind speed at 4-km and hourly resolutions across the contiguous United States (CONUS), based on two future scenarios (RCP4.5 and RCP8.5). This study outlines key steps in developing the high-resolution future solar dataset, including (1) regridding ESM data to a common 20-km resolution grid, (2) correcting ESM biases using the NSRDB, and (3) applying temporal and spatial downscaling methods to generate high-resolution (4-km, hourly) solar projections. Preliminary results indicate that downscaled projections (4-km) captured reasonable spatial patterns when compared to observations across CONUS for four variables. On average across all pixels, 4-km daily-total GHI and DNI projections showed normalized bias (nBias) less than 1% and 6% for GHI and DNI against NSRDB, respectively (nBias less than 1% and 5% for daily-average surface air temperature and surface wind speed). In terms of long-term trend for GHI and DNI, there was no strong increasing or decreasing trend (when compared to surface air temperature), but it showed a very weak decreasing trend.

14 SOLAR ENERGY↗

Weak-form latent space dynamics identification

Recent work in data-driven modeling has demonstrated that a weak formulation of model equations enhances the noise robustness of a wide range of computational methods. In this paper, we demonstrate the power of the weak form to enhance the LaSDI (Latent Space Dynamics Identification) algorithm, a recently developed data-driven reduced order modeling technique. We introduce a weak form-based version WLaSDI (Weak-form Latent Space Dynamics Identification). WLaSDI first compresses data, then projects onto the test functions and learns the local latent space models. Notably, WLaSDI demonstrates significantly enhanced robustness to noise. With WLaSDI, the local latent space is obtained using weak-form equation learning techniques. Compared to the standard sparse identification of nonlinear dynamics (SINDy) used in LaSDI, the variance reduction of the weak form guarantees a robust and precise latent space recovery, hence allowing for a fast, robust, and accurate simulation. We demonstrate the efficacy of WLaSDI vs. LaSDI on several common benchmark examples including viscid and inviscid Burgers', radial advection, and heat conduction. For instance, in the case of 1D inviscid Burgers' simulations with the addition of up to 100% Gaussian white noise, the relative error remains consistently below 6% for WLaSDI, while it can exceed 10,000% for LaSDI. Similarly, for radial advection simulations, the relative errors stay below 15% for WLaSDI, in stark contrast to the potential errors of up to 10,000% with LaSDI. Moreover, speedups of several orders of magnitude can be obtained with WLaSDI. For example applying WLaSDI to 1D Burgers' yields a 140X speedup compared to the corresponding full order model.

97 MATHEMATICS AND COMPUTING↗

Hydrothermal solubility of Dy hydroxide as a function of pH and stability of Dy hydroxyl aqueous complexes from 25 to 250 °C

The rare earth elements (REE) have important applications in green energy technologies. The formation of mineral deposits in geologic systems commonly involves hydrothermal fluids which can mobilize the REE. However, the REE speciation is not well known as a function of pH. The thermodynamic properties of REE hydroxyl complexes used in geochemical models are based on the Helgeson-Kirkham-Flowers (HKF) equation of state parameters which were derived by extrapolation of low temperature experimental and estimated data. In this study, Dy hydroxide solubility experiments are combined with available literature data to improve these models from 25 to 250 °C and optimize the thermodynamic properties of Dy 3+ and Dy hydroxyl complexes using GEMSFITS. Batch-type solubility experiments were conducted from 150 to 250 °C and at saturated water vapor pressure in perchloric acid solutions with initial pH values of 2 to 5 in 0.5 pH unit increments. The measured solubility of Dy hydroxide is retrograde with temperature and decreases with pH. The logarithm of total dissolved Dy molality ranges from –2.3 to –5.3 at 150 °C (pH 4.7–5.5), from –2.4 to –5.6 at 200 °C (pH 3.9–5.1), and from –3.7 to –6.9 at 250 °C (pH of 3.4 and 5.0). The optimized standard partial molal Gibbs energies of formation (Δ f G° T ) derived for Dy 3+ and DyOH 2+ display a close to linear relationship with temperature, fitting with previous optimizations based on DyPO 4 solubility data in the literature. A comparison of the optimized ΔfG°T values for aqueous Dy species with predictions from available HKF parameters indicates significant differences ranging from +11 to –26 kJ/mol between 25 and 250 °C. The experimental fits are used to derive the Dy hydroxide solubility products (K s0 ) and formation constants for the hydrolysis of Dy (β n with n = 1 to 3; Dy 3+ + nOH – = DyOH n 3-n ) as a function of temperature. The optimization method presented yields accurate thermodynamic properties for the Dy 3+ aqua ions and the DyOH 2+ species at the acidic to mildly acidic pH studied whereas more experimental work is needed at near-neutral and alkaline conditions to better constrain the other hydroxyl complexes. Furthermore, the optimized thermodynamic data have a significant impact on geochemical modeling of the mobility and solubility of REE minerals in acidic hydrothermal fluids.

58 GEOSCIENCES↗

Estimating UV-B, UV-Erithemic, and UV-A Irradiances From Global Horizontal Irradiance and MERRA-2 Ozone Column Information

The ground ultraviolet (UV) solar radiation is relevant due to its impacts on plastics degradation (mainly UVA) and on human health (UVB and erithemic UV (UVE)). UV ground measurements are not as ubiquitous as the relatively common global horizontal irradiance (GHI) measurements. Three simple models that estimate the UVA, UVB, and UVE components of solar irradiance from GHI and ozone column information are locally adjusted and validated. Five one-minute datasets from three sites in southeastern South America and two in the United States are used for simultaneous solar irradiance and UV data. All sites correspond to temperate mid-latitude regions. Simultaneous atmospheric total ozone column information is obtained from the reanalysis modern-era retrospective analysis for research and applications (MERRA-2) database for each site. Aside from locally adjusted models, average models with a single set of coefficients are also evaluated. For instance, the best average model is able to estimate UVE with a typical uncertainty below 12% and mean biases between +-3%, relative to the average of the measurements. Similar results are reported for the UVB and UVA components. These results, which can be useful in regions with similar climate and geography, provide a simple way to estimate UV irradiance under all-sky conditions with known uncertainty. This is an alternative to global satellite-based UV estimates, which can have high uncertainties at specific locations. Because MERRA-2 information has a global coverage, when coupled with good satellite-based estimates for GHI, UV irradiances can be estimated by this method over a large territory.

environmental UV radiation↗

RINO: Renormalization Group Invariance with No Labels

A common challenge with supervised machine learning (ML) in high energy physics (HEP) is the reliance on simulations for labeled data, which can often mismodel the underlying collision or detector response. To help mitigate this problem of domain shift, we propose RINO (Renormalization Group Invariance with No Labels), a self-supervised learning approach that can instead pretrain models directly on collision data, learning embeddings invariant to renormalization group flow scales. In this work, we pretrain a transformer-based model on jets originating from quantum chromodynamic (QCD) interactions from the JetClass dataset, emulating real QCD-dominated experimental data, and then finetune on the JetNet dataset -- emulating simulations -- for the task of identifying jets originating from top quark decays. RINO demonstrates improved generalization from the JetNet training data to JetClass data compared to supervised training on JetNet from scratch, demonstrating the potential for RINO pretraining on real collision data followed by fine-tuning on small, high-quality MC datasets, to improve the robustness of ML models in HEP.

Hao, Zichun [Caltech] (ORCID:0000000256244907)↗

Multicycle large-eddy simulations of a direct-injection hydrogen-fueled optical engine

Hydrogen (H 2 ) is a carbon-free chemical energy carrier and one promising solution for achieving effective decarbonization of the transportation sector, particularly for internal combustion engines (ICEs). With a focus on ICEs, and compared to port-fuel injection, direct injection (DI) of gaseous H 2 during the compression stroke offers potential advantages, which include backfire avoidance and reduction of preignition occurrence. In these last two decades, much research, experimental and numerical, has been devoted to understanding H 2 's mixing and combustion processes in ICEs. Computational fluid dynamics modeling efforts commonly rely on unsteady Reynolds-averaged Navier Stokes (URANS) turbulence frameworks, mostly due to their computational affordability. However, many authors have pointed out the opportunity to perform large-eddy simulations (LESs) to investigate the cyclic variability of H 2 engines and assess potential advantages of using LES in place of URANS, especially for lean operation. This study addresses this knowledge gap and presents a computational fluid dynamics (CFD) study of the H 2 DI process in an optical engine operating at relatively low tumble conditions, using multicycle LESs. In conclusion, the manuscript presents a thorough validation of the results against experimental data available from the literature as well as direct comparison with URANS, demonstrating the feasibility of multicycle LESs for CFD modeling of DI H 2 -fueled ICEs.

Direct injection↗

Regional Oil and gas Aerial Methane Synthesis model (ROAMS) v2.0

The Regional Oil and gas Aerial Methane Synthesis model is a tool to convert the results of wide-area, source-resolved aerial methane remote sensing surveys of oil and natural gas infrastructure in a given region into methane emissions inventories (estimates of the magnitude and breakdown of methane emissions from the surveyed infrastructure). The tool leverages databases of source-resolved methane emissions detected in aerial surveys, aerial survey coverage information (which areas were measured and when), data summarizing surveyed oil and natural gas infrastructure and production (derived from third-party databases), as well as state-of-the-art mechanistic emissions simulation tools to characterize emissions too small for the aerial system to see. The regional methane emissions estimates produced by this tool are much more granular in both space and asset type than common satellite- or flux tower-based regional estimates. Unlike other tools for converting site-level measurements into regional emissions estimates, our unique geostatistical approach integrates aerially measured emissions with limited need for statistical extrapolation, which can be highly sensitive to modeler assumptions. As a result, ROAMS-based estimates of regional methane emissions from oil and gas activity are widely viewed as highly credible, as evidenced by the success of Dr. Sherwin's recent paper in Nature.

Sherwin, Evan [Lawrence Berkeley National Laborato↗

The Role of Data Filtering in Open Source Software Ranking and Selection

Faced with more than 100M open source projects, a more manageable small subset is needed for most empirical investigations. More than half of the research papers in leading venues investigated filtering projects by some measure of popularity with explicit or implicit arguments that unpopular projects are not of interest, may not even represent "real" software projects, or that less popular projects are not worthy of study. However, such filtering may have enormous effects on the results of the studies if and precisely because the sought-out response or prediction is in any way related to the filtering criteria.This paper exemplifies the impact of this common practice on research outcomes, specifically how filtering of software projects on GitHub based on inherent characteristics affects the assessment of their popularity. Using a dataset of over 100,000 repositories, we used multiple regression to model the number of stars -a commonly used proxy for popularity- based on factors such as the number of commits, the duration of the project, the number of authors and the number of core developers. Our control model included the entire dataset, while a second filtered model considered only projects with ten or more authors. The results indicated that while certain characteristics of the repository consistently predict popularity, the filtering process significantly alters the relationships between these characteristics and the response. We found that the number of commits exhibited a positive correlation with popularity in the control sample but showed a negative correlation in the filtered sample. These findings highlight the potential biases introduced by data filtering and emphasize the need for careful sample selection in empirical research of mining software repositories. We recommend that empirical work should either analyze complete datasets such as World of Code, or employ stratified random sampling from a complete dataset to ensure that filtering is not biasing the results.

Malviya Thakur, Addi↗

Modeling Spatial Asymmetries in Teleconnected Extreme Temperatures

Abstract Combining strengths from deep learning and extreme value theory can help describe complex relationships between variables where extreme events have significant impacts (e.g., environmental or financial applications). Neural networks learn complicated nonlinear relationships from large datasets under limited parametric assumptions. By definition, the number of occurrences of extreme events is small, which limits the ability of the data-hungry, nonparametric neural network to describe rare events. Inspired by recent extreme cold winter weather events in North America caused by atmospheric blocking, we examine several probabilistic generative models for the entire multivariate probability distribution of daily boreal winter surface air temperature. We propose metrics to measure spatial asymmetries, such as long-range anticorrelated patterns that commonly appear in temperature fields during blocking events. Compared to vine copulas, the statistical standard for multivariate copula modeling, deep learning methods show improved ability to reproduce complicated asymmetries in the spatial distribution of ERA5 temperature reanalysis, including the spatial extent of in-sample extreme events.

Krock, Mitchell L.↗