Search NASA⌕ Search

SEARCH · Search NASA

Results for “Trajectory Data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Techno-Economic Evaluation of Electrified Vehicle Options in Drayage Fleets

The electrification of drayage fleets offers potential economic and operational benefits, but the financial viability of electrified vehicles remains sensitive to battery cost, energy price, and fleet usage patterns. While total cost of ownership (TCO) is a useful benchmark, fleet operators and investors are equally concerned with investment performance metrics such as payback period (PB) and Internal Rate of Return (IRR), which better reflect financial risks and investment return timelines. This study develops a unified techno-economic framework that jointly evaluates TCO, PB, and IRR to determine when electrified trucks become cost-effective alternatives to diesel trucks. Building on a previously developed cost modeling tool and using real-world telematics data from a Class 8 drayage fleet at the Port of Savannah, the analysis incorporates projected battery cost trajectories, electricity and diesel price trends, vehicle efficiency improvements, and multiple battery capacities. Parameter ranges reflect widely cited projections and observed drayage-duty-cycle variability. A surrogate-modeling method approximates economic performance across thousands of battery cost–electricity price combinations, enabling high-resolution identification of conditions that achieve TCO parity, acceptable PB thresholds, and target IRR levels. Additionally, the study estimates the evolving share of the fleet that can feasibly electrify over time under multiple economic metrics. This integrated framework offers a novel, data-driven approach to inform risk-aware decision-making for fleet electrification and supports investment planning under evolving cost and operational conditions.

Sun, Ruixiao [ORNL] (ORCID:0000000341768676)↗

Rapidity-dependent spin decomposition of the nucleon

We revisit the two-dimensional Fourier transform of generalized parton distributions (GPDs) at nonzero skewness. At 𝜂 = 0 it reduces to the standard impact-parameter density, while at 𝜂 ≠ 0 it is an off-forward amplitude that we interpret as a genuine parton–nucleon correlation. Its overall strength (the transverse-plane integral of the density) is fixed by the GPD at the kinematic point 𝑡 =−𝑐 𝜂 =−4⁢𝜂 2 ⁢𝑚$^{2}_{𝑁}$/(1 − 𝜂 2 ) and decreases monotonically with the rapidity gap Δ⁢𝑦 = ln⁡[(1+𝜂)/(1−𝜂)] = 2 artanh⁡(𝜂). This rapidity dependence implies rapidity-modified Ji identities that connect helicity, orbital, and total angular momenta of the correlation in closed form. To quantify these effects, we construct leading-twist quark and gluon GPDs in a string-based conformal framework: conformal moments are parametrized by linear open- and closed-string Regge trajectories with slopes constrained by parton distribution functions (PDFs), hadron/glueball spectroscopy, and form-factor data, and GPDs are reconstructed over the full (𝑥,𝜂,𝑡) domain by Mellin-Barnes inversion with next-to-leading order evolution. We find qualitative agreement (and fair quantitative agreement within quoted uncertainties) for several moments and selected nonsinglet 𝑥-space channels at 𝜇 = 2 GeV when compared with lattice QCD, while we also identify channels with visible tension and discuss likely sources (PDF priors and 𝑡-slope systematics).

Gauge-gravity dualities↗

Impact of Newly Measured 𝛽-Delayed Neutron Emitters around 78 Ni on Light Element Nucleosynthesis in the Neutrino Wind Following a Neutron Star Merger

Neutron emission probabilities and half-lives of 37 𝛽-delayed neutron emitters from 75 Ni to 92 Br were measured at the RIKEN Nishina Center in Japan, including 11 one-neutron and 13 two-neutron emission probabilities and six half-lives for the first time that supersede theoretical estimates. These nuclei lie in the path of the weak 𝑟 process occurring in neutrino-driven winds from the accretion disk formed after the merger of two neutron stars synthesizing elements in the 𝐴∼80 abundance peak. The presence of such elements dominates the accompanying kilonova emission over the first few days and have been identified in the AT2017gfo event, associated to the gravitational wave detection GW170817. Abundance calculations based on over 17,000 simulated trajectories describing the evolution of matter properties in the merger outflows show that the new data lead to an increase of 50%–70% in the abundance of Y, Zr, Nb, and Mo. This enhancement is large compared to the scatter of relative abundances observed in old very metal poor stars and thus is significant in the comparison with other possible astrophysical processes contributing to the light-element production. These results underline the importance of including experimental decay data for very neutron-rich 𝛽 -delayed neutron emitters into 𝑟 -process models.

59 ≤ A ≤ 89↗

Diverse signatures of convergent evolution in cactus-associated yeasts

Many distantly related organisms have convergently evolved traits and lifestyles that enable them to live in similar ecological environments. However, the extent of phenotypic convergence evolving through the same or distinct genetic trajectories remains an open question. Here, we leverage a comprehensive dataset of genomic and phenotypic data from 1,049 yeast species in the subphylum Saccharomycotina (Kingdom Fungi, Phylum Ascomycota) to explore signatures of convergent evolution in cactophilic yeasts, ecological specialists associated with cacti. We inferred that the ecological association of yeasts with cacti arose independently approximately 17 times. Using a machine learning–based approach, we further found that cactophily can be predicted with 76% accuracy from both functional genomic and phenotypic data. The most informative feature for predicting cactophily was thermotolerance, which we found to be likely associated with altered evolutionary rates of genes impacting the cell envelope in several cactophilic lineages. We also identified horizontal gene transfer and duplication events of plant cell wall–degrading enzymes in distantly related cactophilic clades, suggesting that putatively adaptive traits evolved independently through disparate molecular mechanisms. Notably, we found that multiple cactophilic species and their close relatives have been reported as emerging human opportunistic pathogens, suggesting that the cactophilic lifestyle—and perhaps more generally lifestyles favoring thermotolerance—might preadapt yeasts to cause human disease. This work underscores the potential of a multifaceted approach involving high-throughput genomic and phenotypic data to shed light onto ecological adaptation and highlights how convergent evolution to wild environments could facilitate the transition to human pathogenicity.

59 BASIC BIOLOGICAL SCIENCES↗

Data-driven discovery of dynamics from time-resolved coherent scattering

Coherent X-ray scattering (CXS) techniques are capable of interrogating dynamics of nano- to mesoscale materials systems at time scales spanning several orders of magnitude. However, obtaining accurate theoretical descriptions of complex dynamics is often limited by one or more factors—the ability to visualize dynamics in real space, computational cost of high-fidelity simulations, and effectiveness of approximate or phenomenological models. In this work, we develop a data-driven framework to uncover mechanistic models of dynamics directly from time-resolved CXS measurements without solving the phase reconstruction problem for the entire time series of diffraction patterns. Our approach uses neural differential equations to parameterize unknown real-space dynamics and implements a computational scattering forward model to relate real-space predictions to reciprocal-space observations. This method is shown to recover the dynamics of several computational model systems under various simulated conditions of measurement resolution and noise. Moreover, the trained model enables estimation of long-term dynamics well beyond the maximum observation time, which can be used to inform and refine experimental parameters in practice. Finally, we demonstrate an experimental proof-of-concept by applying our framework to recover the probe trajectory from a ptychographic scan. Our proposed framework bridges the wide existing gap between approximate models and complex data.

36 MATERIALS SCIENCE↗

Second-generation downscaled earth system model data using generative machine learning

The second-generation Sup3rCC dataset provides high-resolution meteorological data generated through the downscaling of multiple earth system models (ESMs) from the Coupled Model Intercomparison Project Phase 6 (CMIP6). This downscaling is performed through application of a generative machine learning approach called Super-Resolution for Renewable Resource Data (sup3r). This dataset builds on the first-generation Sup3rCC data by applying improved bias correction methods and adding downscaled precipitation to the output variables. As with the first Sup3rCC version, the data still include temperature, wind speed and direction at multiple heights, pressure, three components of downwelling solar radiation, and relative humidity—all at 4-kilometer (km) hourly resolution over the contiguous United States. This is a 25x spatial enhancement and 24x temporal enhancement of the source 100-km daily-average ESM data. This extension of the Sup3rCC dataset includes data from six ESMs from two shared socioeconomic pathways (SSPs) totaling 400 years of data with multiple future projections of changing meteorological conditions. The scenario selection was based on a structured evaluation of historical ESM skill and comprehensive representation of possible trajectories of future climate change in temperature, humidity, precipitation, solar irradiance, and near-surface wind speeds. The inclusion of multiple future projections is intended to enable users to assess key drivers of un 36 certainty and variability. All data are double-bias corrected, resulting in a product that can be used out-of-the-box for energy system analysis with minimal historical bias. The potential applications of Sup3rCC data extend to various topics in renewable energy resource assessment, energy systems modeling, and grid resilience studies. High-resolution future meteorological projections are critical for evaluating the effects of changing meteorological conditions on renewable energy generation, energy demand, and for optimizing energy storage and grid infrastructure. The 4-km hourly resolution of the downscaled data enables understanding of spatial and temporal variability at the scales necessary for energy system operational planning. In addition, the dataset can support risk assessments by providing detailed information on possible future extreme weather events and long-term meteorological variability at scales relevant to energy infrastructure. By offering an enhanced representation of possible future meteorological conditions, the second-generation Sup3rCC dataset enables more precise modeling of energy resilience and adaptation strategies in response to changing meteorological conditions.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Phosphate amendment drives bloom of RNA viruses after soil wet-up

Soil rewetting after a dry period results in a surge of activity and succession in both microbial and DNA virus communities. Less is known about the response of RNA viruses to soil rewetting—while they are highly diverse and widely distributed in soil, they remain understudied. We hypothesized that RNA viruses would show temporal succession following rewetting and that phosphate amendment would influence their trajectory, as viral proliferation may cause phosphorus limitation. Using 39 time-resolved metatranscriptomes and amplicon data, 2190 RNA viral populations were identified across five phyla, with 26 % of these predicted to infect bacteria, and 11 % fungi. Only 1.2 % of viral populations had annotated capsid genes, suggesting most persist via intracellular replication without a free virion phase. Phosphate amendment altered RNA viral community composition within the first week and amended vs. unamended communities remained distinguishable for up to three weeks. While the overall host community remained stable, certain bacterial populations showed reduced abundance in phosphate-amended soils, likely due to increased viral lysis, as RNA bacteriophages proliferated significantly. Notably, 60 % of the viruses with increased abundance under phosphate amendment belonged to basal Lenarviricota clades rather than well-known groups like Leviviricetes. We estimate RNA bacteriophage infections may affect 10 7 –10 9 bacteria per gram of soil, aligning with the total bacterial population (10 7 –10 10 g -1 soil), suggesting that RNA phages significantly influence bacterial communities post-wet-up, with phosphorus availability modulating this effect.

59 BASIC BIOLOGICAL SCIENCES↗

Poster Abstract: Leveraging Large Language Models to Reveal Interpretable Cooling Behaviors from Smart Thermostat Data

Frequent heatwaves and hot summers increasingly challenge occupant comfort, health, and energy grid stability. Addressing these challenges requires a detailed understanding of household cooling behaviors, such as thermostat adjustments and adaptive responses to extreme conditions. Traditional analyses often rely on aggregated numerical metrics that overlook subtle but important household-specific variations. In this study, we introduce a generalizable methodology that integrates large language models (LLMs) with vision capabilities to enable scalable and detailed analysis of residential thermostat data. Using Ecobee's Donate Your Data (DYD) dataset—which provides five-minute records of indoor temperatures, thermostat setpoints, and HVAC runtimes—we focus on two U.S. cities with contrasting summer climates : Austin (TX) and Phoenix (AZ). Because raw time-series data are not well suited for direct LLM analysis, we transform them into visual representations, such as daily indoor temperature trajectories and weekly runtime histograms, to better capture behavioral variations. Leveraging LLMs' visual interpretation, we extract descriptive behavioral features, including temperature preferences, time-of-day cooling orientation, anticipatory versus reactive heatwave responses, and behavioral consistency. These semantic features support unsupervised clustering to identify distinct occupant archetypes at scale, revealing differences—such as morning-centric anticipatory coolers versus households that shift toward warmer setpoints during heatwaves—that can inform demand response, resilience planning, and health-aware interventions. By converting raw numerical data into interpretable behavioral patterns, this methodology enables scalable and practical analysis of occupant behavior, supporting actionable insights for comfort, resilience, and energy management.

Nihar, Kopal↗

Similarity Metric for Data Optimization and Efficient Training of Reactive Machine Learning Force Fields for Hydrocarbon Radiolysis

Radiolysis is a common approach to sterilize polymers, chemically modify them for upcycling, and accelerate their decomposition for recycling purposes. Reactive molecular dynamics (MD) simulations provide a powerful tool to generate atomic-level trajectories of the reactive processes and quantify radiolytic chemical degradation pathways. For this, machine learning (ML) surrogate models for reactive force fields with quantum mechanical accuracy are now widely used, which require ML training data sets that can provide information on atomic environments for target chemical systems. However, radiolysis chemistry can be highly complex and diverse, which poses significant challenges for generating training data to parametrize ML models. In this regard, we developed a method for optimizing the training data set using a cosine similarity metric to help guide training set selection for radiolysis of polyethylene, a model hydrocarbon polymer, as well as to enhance the transferability of our reactive ML force field (MLFF) to a variety of molecular and polymeric systems. Our approach performs atom-by-atom comparisons between local atomic environments to pinpoint important data points associated with rare and localized events, such as radiolysis damage within structures. We apply this approach to train the Chebyshev Interaction Model for Efficient Simulation (ChIMES) MLFF model, which expresses the atomic interaction potentials in terms of linear combinations of many-body Chebyshev polynomials. We first show that our method can reduce our training set size by ∼70% while improving overall accuracy compared to more standard MD model fitting approaches. We then validate our optimum model against diverse hydrocarbon simulation data, including simple alkanes and systems with unsaturated carbon bonds, over a wide range of thermodynamic conditions. Finally, we use our ChIMES model to perform MD simulations of radiolytic damage with large-scale systems that help avoid system size effects. Overall, our approach yields an MD force field that retains most of the accuracy of the underlying quantum method while yielding many orders of improvement in computational efficiency. In conclusion, our efforts will have impact on future hydrocarbon polymer radiolysis studies, where the chemical details of the polymer–radiation interactions can have a strong effect on the resulting products observed in experiments.

Hydrocarbons↗

Utah FORGE: 2024 Discrete Fracture Network Model Data

The Utah FORGE 2024 Discrete Fracture Network (DFN) Model dataset provides a set of files representing discrete fracture network modeling for the FORGE site near Milford, Utah. The dataset includes four distinct DFN model file sets, each corresponding to different time frames and modeling approaches in 2024. These models characterize both natural and induced fractures in the geothermal reservoir, which consists of crystalline granitic and metamorphic rock approximately 8,000 feet below the ground surface. The dataset includes a reference DFN model from February 2024 that incorporates planar fractures and well trajectories, as well as upscaled permeability, porosity, compressibility, and storage values on specified grids. Additionally, there are models based on new microseismic (MEQ) data from May and July 2024, including fracture planes fitted to the latest MEQ catalog datasets, tensile fractures from hydraulic stimulation, and an alternative connected DFN for modeling purposes. Coordinate data is provided in both global and local frames, with detailed instructions on the transformations used to align with principal stress orientations. The dataset also includes notes and calculation files for estimating fracture sizes and differences between various fracture sets. There are subfolders for Global Coordinates and Local Coordinates. To move from the global to the local coordinate frame, fractures and wells were a) rotated 20 degrees counterclockwise looking down about the global point (335376.400482041, 4263189.99998761, 250.093546450195) to better align with the principal stresses; and b) translated by (-335408.68, -4263010.9, 1150). Upscaled permeability values using the _XYZ suffix show directions with respect to the global XYZ coordinate frame, while those using the _IJK suffix are aligned with local coordinate frame.

15 GEOTHERMAL ENERGY↗

Score-based deterministic density sampling

We propose a deterministic sampling framework using Score-Based Transport Modeling for sampling an unnormalized target density π given only its score ∇ log π. Our method approximates the Wasserstein gradient flow on KL($f_t$∥π) by learning the time-varying score ∇ log $f_t$ on the fly using score matching. While having the same marginal distribution as Langevin dynamics, our method produces smooth deterministic trajectories, resulting in monotone noise-free convergence. We prove that our method dissipates relative entropy at the same rate as the exact gradient flow, provided sufficient training. Numerical experiments validate our theoretical findings: our method converges at the optimal rate, has smooth trajectories, and is often more sample efficient than its stochastic counterpart. Experiments on high-dimensional image data show that our method produces high-quality generations in as few as 15 steps and exhibits natural exploratory behavior. The memory and runtime scale linearly in the sample size.

97 MATHEMATICS AND COMPUTING↗

Digital twin framework for PIP-II linac: AI-driven multi-scale modeling from ion source to 800 MeV

The PIP-II superconducting linac at Fermilab is designed to deliver multi-megawatt proton beams for neutrino physics and other high-intensity applications. To expedite commissioning and enhance operational reliability, we have developed an EPICS-based data flow framework that seamlessly integrates digital twins (DT) with physical twins (PT). These digital twins comprise high-fidelity beam dynamics models or data-driven surrogate models connected to their physical counterparts through real-time diagnostics and advanced machine-learning algorithms.Central to this framework is Linac_Gen, an accelerated simulation tool that incorporates convolutional neural networks, random forests, and genetic algorithms to provide up to a tenfold speedup in optimizing the accelerator geometry model. An EPICS translator layer ensures interoperability by efficiently mapping lattice parameters across diverse simulation platforms.Our EPICS-based framework supports multiple operational modes—monitoring, passive learning, closed-loop control, and online learning—covering the entire machine lifecycle. By leveraging HPC resources and multi-objective optimization techniques, the digital twin enables adaptive trajectory correction, real-time fault detection, and predictive modeling of beam stability. This comprehensive approach paves the way for robust, high-intensity operation and data-driven accelerator R&D at Fermilab.

Pathak, Abhishek [Fermilab]↗

Extracting Vehicle Trajectories from Partially Overlapping Roadside Radar

This work presents a methodology for extracting vehicle trajectories from six partially-overlapping roadside radars through a signalized corridor. The methodology incorporates radar calibration, transformation to the Frenet space, Kalman filtering, short-term prediction, lane-classification, trajectory association, and a covariance intersection-based approach to track fusion. The resulting dataset contains 79,000 fused radar trajectories over a 26-h period, capturing diverse driving scenarios including signalized intersections, merging behavior, and a wide range of speeds. Compared to popular trajectory datasets such as NGSIM and highD, this dataset offers extended temporal coverage, a large number of vehicles, and varied driving conditions. The filtered leader–follower pairs from the dataset provide a substantial number of trajectories suitable for car-following model calibration. The framework and dataset presented in this work has the potential to be leveraged broadly in the study of advanced traffic management systems, autonomous vehicle decision-making, and traffic research.

33 ADVANCED PROPULSION SYSTEMS↗

Quarterly Soil Core and Root Analyses from the Missouri Ozarks AmeriFlux (MOFLUX) Site, Ashland, Missouri, 2017-2023

This dataset contains quarterly soil core measurements from the Missouri Ozarks AmeriFlux (MOFLUX) site located at the University of Missouri’s Thomas H. Baskett Wildlife Research and Education Area near Ashland, Missouri. These data will be used to parameterize an ensemble of MOFLUX-optimized soil carbon-nitrogen models, used to simulate carbon (C) and nitrogen (N) cycling responses to future hydroclimatic scenarios and the trajectory of soil C stocks with concomitant forest decline. Beginning in 2017, eight soil cores were collected approximately quarterly near plot 1 of the southeast transect, near the automated soil respiration flux chambers, from 0–15 cm depth. Data are currently available through 2023 (2017-06-14 to 2023-11-13); additional observations will be appended to this dataset as they become available. Cores were analyzed for gravimetric moisture content, pH, total carbon and nitrogen, texture, microbial biomass carbon and nitrogen, and extractable dissolved organic carbon and nitrogen. This dataset contains one data file in comma separate (*.csv) format. Additional metadata are provided: one data dictionary and a file-level metadata file in comma separate (*.csv) format and a user guide in PDF (*.pdf) format.

54 ENVIRONMENTAL SCIENCES↗

Spatiotemporal Dynamics of the Relative Abundance of Soil Nutrient‐Degrading Enzyme‐Encoding Genes Across Continental US Ecoregions

Understanding the spatiotemporal patterns in the relative abundance of soil extracellular enzyme‐encoding genes is critical for predicting microbial responses to environmental change and their potential role in nutrient cycling. Yet, integrating novel metagenomic observations with spatiotemporal environmental gradients to infer regional patterns and future trajectories has remained unclear. To address this gap, we applied a machine learning (ML) approach, integrating soil metagenomic data with environmental variables—soil properties, topography, vegetation, and climate—to predict the relative abundance of enzyme‐encoding genes for soil carbon (C), nitrogen (N), and phosphorus (P) across surface soils of the continental United States. We assessed potential responses under future emission scenarios (SSP2‐4.5 and SSP5‐8.5) by comparing a baseline (1985–2014) to a future period (2071–2100). The ML model explained 57%–63% of baseline variation. Precipitation was identified as the most influential factor for the relative abundance of C‐ and N‐degrading enzyme‐encoding genes, while slope length, representing horizontal distance that water can travel downslope, was the primary driver for P‐degrading enzyme‐encoding genes abundance. Projections revealed spatially heterogeneous shifts across continental US ecoregions: the relative abundance of C‐ and N‐degrading enzyme‐encoding genes decreased in wetter ecoregions and increased in drier ecoregions under future climate, while P‐degrading enzyme‐encoding genes abundance decreased significantly in semiarid and Mediterranean ecoregions. This study demonstrates the utility of metagenomic data for mapping soil genetic potential and predicting its regional response to environmental change, to inform ecosystem management strategies.

extracellular enzyme-encoding genes↗

A Data-Driven Approach for High-Impedance Fault Localization in Distribution Systems

Accurate and quick identification of high-impedance faults (HIFs) is critical for the reliable operation of distribution systems. Unlike other faults in power grids, HIFs are very difficult to detect by conventional overcurrent relays due to the low fault current. Although HIFs can be affected by various factors, the voltage-current characteristics can substantially imply how the system responds to the disturbance and thus provides opportunities to effectively localize HIFs. In this work, we propose a data-driven approach for the identification of HIF events. To tackle the nonlinearity of the voltage-current trajectory, first, we formulate optimization problems to approximate the trajectory with piecewise functions. Then we collect the function features of all segments as inputs and use the support vector machine approach to efficiently identify HIFs at different locations. Numerical studies on the IEEE 123-node test feeder demonstrate the validity and accuracy of the proposed approach for real-time HIF identification.

explainable artificial intelligence↗

StOKeDMD: Streaming Occupation kernel dynamic mode decomposition

Dynamic mode decomposition (DMD) has become a common technique for constructing surrogate models for dynamical systems from observed system states. The Occupation Kernel DMD (OKDMD) method proposed in (Rosenfeld et al., 2022) and (Rosenfeld et al., 2024) is a Liouville operator based method that builds surrogate models from system state trajectories. Here, this paper proposes an extension of OKDMD to the case when the system states are observed in a streaming fashion, i.e., only a small fraction of the state trajectory is available at a given time. The developed method, Streaming Occupation Kernel DMD (StOKeDMD), accommodates the streaming data input by leveraging properties of specific choices of kernel functions and occupation kernels. We apply the StoKeDMD method as a compression method for streaming data, analyze the memory complexity, and demonstrate the performance of StoKeDMD in the compression of streaming data generated from a Lorenz system and a fluid flow simulation.

97 MATHEMATICS AND COMPUTING↗

A Data-Driven Approach for High-Impedance Fault Localization in Distribution Systems: Preprint

Accurate and quick identification of high-impedance faults (HIFs) is critical for the reliable operation of distribution systems. Unlike other faults in power grids, HIFs are very difficult to detect by conventional overcurrent relays due to the low fault current. Although HIFs can be affected by various factors, the voltage-current characteristics can substantially imply how the system responds to the disturbance and thus provides opportunities to effectively localize HIFs. In this work, we propose a data-driven approach for the identification of HIF events. To tackle the nonlinearity of the voltage-current trajectory, first, we formulate optimization problems to approximate the trajectory with piecewise functions. Then we collect the function features of all segments as inputs and use the support vector machine approach to efficiently identify HIFs at different locations. Numerical studies on the IEEE 123-node test feeder demonstrate the validity and accuracy of the proposed approach for real-time HIF identification.

explainable artificial intelligence↗