Search NASA⌕ Search

SEARCH · Search NASA

Results for “data exchange”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 235 records · Page 13

A universal implementation of radiative effects in neutrino event generators

Due to the similarities between electron-nucleus (eA) and neutrino-nucleus scattering (νA), eA data can contribute key information to improve cross-section modeling in eA and hence in νA event generators. However, to compare data and generated events, either the data must be radiatively corrected or radiative effects need to be included in the event generators. We implemented a universal radiative corrections program that can be used with all reaction mechanisms and any eA event generator. Our program includes real photon radiation by the incident and scattered electrons, and virtual photon exchange and photon vacuum polarization diagrams. It uses the “extended peaking” approximation for electron radiation and neglects charged hadron radiation. This method, validated with GENIE, can also be extended to simulate νA radiative effects. This work facilitates data-event-generator comparisons used to improve νA event generators for the next-generation of neutrino experiments. Program Title: emMCRadCorr CPC Library link to program files:https://doi.org/10.17632/hmsxg82vnf.1 Developer's repository link:https://github.com/e4nu/emMCRadCorr Licensing provisions: AGPLv3 Programming language:C++ Nature of problem: Radiative effects can significantly modify the event kinematics and the resulting cross-sections. Such effects must be accounted for when comparing event generators to eA data. Existing radiative correction codes are tailored to specific processes and topologies, and are limited to a restricted phase space defined by the spectrometer acceptance. Therefore, a more general approach is required to apply radiative corrections to semi-inclusive and exclusive eA measurements. Solution method: Our program incorporates real photon radiation from both the incident and scattered electrons, as well as virtual photon exchange and photon vacuum polarization effects. It employs the “extended peaking” approximation for electron radiation while neglecting contributions from charged hadron radiation. The code is fully decoupled from event generator codes and can be used for all event generators in the market.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Data for Resilience and Reliability [Slides]

The Energy to Communities (E2C) peer-learning cohort program provides technical assistance to groups of 15 community entities around a common energy topic over the course of 6 months. Every month, participants join a virtual meeting where they hear from experts and exchange strategies and best practices with their peers. This cohort, "Planning for Major Energy Disruptions in the Southeast" focuses on strategies for entities in the Southeast to quickly recover from impacts to energy infrastructure caused by extreme events, such as severe weather and natural disasters. This presentation focuses on the data available related to energy disruptions. The workshop is on April 23, 2026.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Public Reference Data for Megawatt-Scale Hydrogen Electrolysis - Simulated Wave

The U.S. Department of Energy and the National Laboratory of the Rockies (NLR) demonstrate hydrogen electrolysis, hydrogen compression and storage, and variable hydrogen fuel cell power production using megawatt-scale equipment at NLR’s Flatirons Campus as part of the Advanced Research on Integrated Energy Systems (ARIES) initiative. This dataset represents part of that effort and is intended for academic, national laboratory, industrial, and other stakeholders to plan, design, and validate models of megawatt-scale hydrogen technologies and diverse energy infrastructure nationwide. These data provide a baseline for how existing hydrogen electrolysis technologies perform when coupled with various energy technologies. Future datasets will demonstrate how existing hydrogen fuel cell technologies can provide controllable, dispatchable, and variable power output for artificial intelligence (AI) data centers and other variable loads. This dataset entry describes hydrogen production using a single, simulated wave energy conversion device. The electrolyzer is a 1.25-MW proton exchange membrane type MC250 system manufactured by Nel Hydrogen. While the unit supports up to 2.5 MW of electrolysis, NLR only has a single 1.25-MW electrolysis stack. For the wave energy, NLR used a wave energy converter model from PacWave. These devices can be equipped with accumulators and pressure relief values to smooth the power output by storing and releasing hydraulic energy. Using a peak power output of 10 MW, the model created two 25-minute profiles: one with and one without the accumulators and pressure relief valves. To down select the profile data from the native resolution of 20 Hz to 1 Hz, NLR took the mean of every 20 data points. NLR experimented with two simulated wave energy power plants: one that peaks at 10 MW, and one that peaks at 5 MW. These profiles were scaled for the physical 1.25 MW electrolyzer by multiplying the original profiles by one eighth and one quarter, respectively. The first profile matches the capacity rating of eight of the 1.25 MW electrolyzers, while the second matches four electrolyzers. Finally, NLR experimented with two settings for the electrolyzer power supply minimum and maximum current ramp rates (gain and slew): 200 and 400 amperes per second. The simulated profiles were translated from power (kilowatts) to current (amperes) using a curve fit with calibration data and sent to the electrolyzer power supply at 1-Hz frequency. These datasets report relevant hydrogen balance-of-plant and system data, all captured at 1 Hz, including hydrogen mass production measured with an Emerson Coriolis flow meter. Each .zip file represents a single wave electrolysis experiment and is formatted as follows: {technology}-{accumulator?}_{number of 1.25 MW electrolyzers connected}-{electrolyzer ramp rate in amperes/second} For instance, “wavePacWave-Noacc_4-400.zip” represents the 25 minute-long experiment using the PacWave’s wave energy converter model, equipped with no accumulator, connected to four 1.25-MW electrolyzers with their power supplies set to a maximum current ramp rate (gain and slew) of 400 A/s. Each .zip folder contains the following files: A .csv file containing raw data. An .xlsx file explaining all the fields in the raw data. A .png plot showing the time series of hydrogen production in kilograms per hour, electrolysis power consumption, and input wave power. An experiment, labeled “characterization_200.zip”, demonstrates the MC250 electrolyzer steady-state response with 30 minute load steps for a total duration of 5 hours. Finally, a .csv file is provided with all wave profiles combined into one dataset labeled "combined_wave_experiments.csv". NLR also built an AI/machine-learning predictive model based on these datasets. The model ingests the electrolyzer current command in amperes, as well as various pressures and temperatures across the system, and predicts hydrogen output in kilograms per hour. The complete model can be found at https://huggingface.co/NatLabRockies/ptmelt-hydrogen-electrolysis.

08 HYDROGEN↗

Helium-4 gravitational form factors: Exchange currents

We evaluate the leading exchange corrections to the helium-4 gravitational form factors (GFFs) to momenta of the order of the nucleon mass. We use both the K-harmonic method with simple pair nucleon potential, and a Jastrow trial function using the Argonne 𝑣 14 potential, to evaluate the helium-4 GFFs. The exchange current contributions include the pair interaction, plus the seagull and the pion exchange interactions, modulo the recoil corrections. To estimate the off-shellness of the pion nucleon coupling in this momenta range, we discuss the results using either the pseudoscalar (PS) or pseudovector (PV) pion-nucleon couplings. When the PV coupling is used, the pair diagram contribution is higher order in the nonrelativistic expansion. The results for the helium-4 A-GFF are comparable to those given by the impulse approximation, especially for the PS coupling using both the K-harmonic method and variational method. The exchange current contributions with the PS coupling for the charge form factor of helium-4, yield better agreement with the existing data over a broad range of momenta, especially when the Argonne 𝑣 14 potential including the D-wave admixture is used.

A ≤ 5↗

Evapotranspiration partitioning estimates from 8 methods from 47 NEON sites, 2019-2021

This dataset provides daily estimates of evapotranspiration (ET) and the transpiration-to-evapotranspiration ratio (T/ET) across 47 terrestrial National Ecological Observatory Network (NEON) sites spanning diverse environmental and biome conditions in the United States across three years of data (2019-2021). Daily ET is reported in both energy units (MJ m⁻² day⁻¹) and equivalent water depth (mm day⁻¹), assuming a constant latent heat of vaporization of 2.45 MJ/kg. The primary method uses a hybrid recurrent neural network–Penman–Monteith framework (RNN-PM), which integrates physically based surface energy balance constraints with data-driven learning to partition ET into transpiration and evaporation components. Model inputs include in situ meteorological observations (air temperature, vapor pressure deficit, wind speed, and radiation) combined with satellite-derived land surface temperature, leaf area index, and soil moisture. For benchmarking and uncertainty assessment, T/ET estimates from seven additional models are included: Priestley-Taylor Jet Propulsion Laboratory (PT-JPL), Penman-Monteith (P-M), Two-Source Energy Balance (TSEB), Support Vector Regression (SVR), and Categorical Boosting (CatBoost), among others—spanning empirical, machine-learning, and process-based approaches (see methods section or linked publication for detailed descriptions). Data Package Contents: The dataset a csv files containing daily ET and T/ET estimates for each site and model, along with associated metadata files these variables. Data can be accessed using common spreadsheet software (e.g., Microsoft Excel, LibreOffice) or programming environments such as R or Python. Together, these data support cross-site comparisons of ecosystem water use, evaluation of ET partitioning methods, and development of improved land–atmosphere exchange models.

EARTH SCIENCE > ATMOSPHERE↗

Technoeconomic Analysis Round Robin of a Retrofit of the Ivanpah Concentrating Solar Plant with a Molten-Salt System with Thermal Energy Storage

While the fidelity of technoeconomic analysis (TEA) models for concentrating solar thermal systems has improved in recent years, there is a lack of consensus on the specific inputs used to forecast performance of a newly built tower system due to a lack of validations and post-mortems available to the public. This effort is a joint initiative between multiple international organizations to validate and compare their TEA models. The specific case study is a proposed retrofit of one unit of the Ivanpah Solar Energy Generating System to include molten-salt storage, replacing the steam generation system with a molten-salt receiver, salt-to-steam heat exchanger train, and new balance of plant while keeping the existing steam turbine, solar field, and interconnection in place, using plant data for calibration. This manuscript discusses several of the agreed-upon assumptions for this study as well as a preliminary analysis from prior work that motivates the study.

14 SOLAR ENERGY↗

Public Reference Data for Megawatt-Scale Hydrogen Electrolysis – Simulated Wind

The U.S. Department of Energy and the National Laboratory of the Rockies (NLR) demonstrate hydrogen electrolysis, hydrogen compression and storage, and variable hydrogen fuel cell power production using megawatt-scale equipment at NLR’s Flatirons Campus as part of the Advanced Research on Integrated Energy Systems (ARIES) initiative. This dataset represents part of that effort and is intended for academic, national laboratory, industrial, and other stakeholders to plan, design, and validate models of megawatt-scale hydrogen technologies and diverse energy infrastructure nationwide. These data provide a baseline for how existing hydrogen electrolysis technologies perform when coupled with various energy technologies. Future datasets will demonstrate how existing hydrogen fuel cell technologies can provide controllable, dispatchable, and variable power output for artificial intelligence (AI) data centers and other variable loads. This dataset entry describes hydrogen production using a single, simulated wind turbine. The electrolyzer is a 1.25-MW proton exchange membrane type MC250 system manufactured by Nel Hydrogen . While the unit supports up to 2.5 MW of electrolysis, NLR only has a single 1.25-MW electrolysis stack. For the simulated wind energy profiles, NLR used OpenFAST to simulate a 3.4-MW International Energy Agency (IEA) reference wind turbine. The hour-long wind energy profiles varied over wind turbulence intensity (Class A or Class C) and average wind speed (5, 7, or 9 m/s). To match the power limits of the 1.25-MW electrolyzer and 3.4-MW IEA wind turbine most effectively and to maximize the efficiency of hydrogen production at a given average wind speed, the profiles were sometimes scaled by two times. This means that, in some cases, the experimental setup assumed two 1.25-MW electrolyzers were coupled with the wind turbine, representing a total maximum electrolysis load of 2.5 MW. Finally, NLR experimented with two settings for the electrolyzer power supply minimum and maximum current ramp rates (gain and slew): 200 and 400 amperes per second. The simulated profiles were translated from power (kilowatts) to current (amperes) using a curve fit with calibration data and sent to the electrolyzer power supply at 1-Hz frequency. These datasets report relevant hydrogen balance-of-plant and system data, all captured at 1 Hz, including hydrogen mass production measured with an Emerson Coriolis flow meter. Each .zip file represents a single wind turbine electrolysis experiment and is formatted as follows: {technology}-{average wind speed}-{turbulence class}_{number of 1.25 MW electrolyzers connected}-{electrolyzer ramp rate in amperes/second} For instance, “windIEA3.4-5ms-C_2-400.zip” represents the hour-long experiment using the IEA 3.4-MW turbine, subjected to an average wind speed of 5 m/s and Class C wind turbulence, and connected to two 1.25-MW electrolyzers with the power supply set to a maximum current ramp rate (gain and slew) of 400 A/s. Each .zip folder contains the following files: A .csv file containing raw data. An .xlsx file explaining all the fields in the raw data. A .png plot showing the time series of hydrogen production in kilograms per hour, electrolysis power consumption, and input wind turbine power. An experiment labeled “characterization_200.zip” demonstrates the MC250 electrolyzer steady-state response with 30 minute load steps for a total duration of 5 hours. Finally, a .csv file is provided with all simulated wind experiments combined into one dataset labeled "combined_wind_experiments.csv". NLR also built an AI/machine-learning predictive model based on these datasets. The model ingests the electrolyzer current command in amperes, as well as various pressures and temperatures across the system, and predicts hydrogen output in kilograms per hour. The complete model can be found at https://huggingface.co/NatLabRockies/ptmelt-hydrogen-electrolysis .

08 HYDROGEN↗

Diagenesis is key to unlocking outcrop fracture data suitable for quantitative extrapolation to geothermal targets

Exceptionally large, well-exposed sandstone outcrops in New York provide insights into folds, deformation bands, and fractures that could influence permeability, heat exchange, and stimulation outcomes of geothermal reservoir targets. Cambrian Potsdam Sandstone with <5% porosity contains decimeter-scale open, angular-limbed monoclines <0.5 km apart with associated low-porosity mm-wide cataclastic deformation bands. Crossing and abutting relationships among sub-vertical opening-mode fractures show four chronological Sets A–D, striking NNW, NE, NW, and ENE, respectively. Fracture lengths and heights range from millimeters to tens of meters. Sets A and C macro-fractures, and possibly B and D, contain quartz deposits. All sets have abundant associated quartz cemented microfractures that also record set orientations and crosscutting relations. Quartz cement deposits—evidence of diagenesis—are the key to identifying attributes of outcrop fractures suitable for extrapolation to geothermal targets in sandstones because they show which fractures formed in the subsurface. Set A fluid inclusion homogenization temperatures (120°C–129°C) are compatible with fracture at >3 km depth. Fractures are stiff and those ≥0.05 mm (Set C) and ≥0.1 mm (Set A) are open and potentially conducive to flow. Sets A and D are abundant in outcrops with close fracture spacing—0.18 m and 0.68 m, respectively—and define a rectangular connectivity network dominated by crossing and abutting X and Y nodes. Set A aperture distributions follow a power law with slope –0.8 up to 0.15 mm; other sets have lognormal distributions. Set A and D microfractures are weakly clustered, while macro-fractures commonly have 1D anticlustered (regular or periodic) arrangements at shorter length scales (<0.2 m). Sub-horizontal fractures are barren and may have formed near the surface. Fracture heights, lengths, and spatial arrangements show good trace connectivity but low open connectivity. For geothermal applications, outcrop results predict low initial well-test permeabilities owing to quartz disconnecting open fractures, but stimulation of closely spaced microfractures and partly open macro-fractures could yield high surface area for heat exchange. Quantitative extrapolation of key fracture attributes like abundance, orientation, spatial arrangement, length, and open fracture connectivity is possible from outcrops to fractured reservoirs if differing thermal histories and diagenesis are accounted for.

02 PETROLEUM↗

Ion Distribution and Cation Exchange at Mica–Electrolyte Interfaces Probed with Deep Potential Molecular Dynamics

Here, we investigate the Stern layer structure and cation exchange mechanism at muscovite mica-electrolyte interfaces using nanosecond timescale molecular dynamics simulations based on deep neural network interatomic potentials trained on Density Functional Theory (DFT) data. Focusing on mica with exposed surface K + interfaced with aqueous NaCl and mica with surface Na + interfaced with KCl solution, we find that K + remains predominantly in inner-sphere configurations, while Na + exhibits notable populations in outer-sphere states. Most importantly, our simulations show that contact with an electrolyte solution results in the co-adsorption of multiple cation species, making the mica surface locally overcharged and thus reshaping the cation speciation in a manner that enhances the tendency of neighboring surface cations to desorb. These findings are consistent with recent experimental observations that co-adsorption of different cation species induces changes in cation speciation and slow kinetics of cation exchange at the muscovite-water interface, providing a basis for their detailed understanding.

36 MATERIALS SCIENCE↗

Energy performance of an operational government building retrofitted with ceiling phase change material tiles in a mixed-humid climate

The aging U.S. building stock requires various retrofit measures to enhance their energy efficiency. Here, this study explores the integration of thermal energy storage and advanced building controls as viable retrofit solutions for load flexibility and peak demand response while maintaining the occupants' comfort. A detailed assessment is conducted on the energy use of an administrative building in Sumner County, Kansas, focusing on the implementation of phase change materials (PCMs) in the ceiling of occupied zones. First, a time-resolved, whole-building energy model is developed in EnergyPlus, incorporating complex thermal behaviors such as air exchange between the plenum space and occupied zones, envelope leakage, and operational schedules. The model is then validated using experimental field test data, and subsequently a parametric assessment of key PCM properties and application strategies is performed to evaluate cooling electricity demand benefits. The parametric study shows that the optimal retrofit strategy, comprising a PCM with 23°C peak melting temperature, 0.125 in. (3.17 mm) thickness, and 150 kJ/kg latent heat, combined with active controls that include 8 h of precooling, forced convection under the ceiling, and a 2°C thermostat setback during peak hours, can result in a maximum load shift during the peak period of 99.6 % and the total electricity savings during the peak period of 98.9 % for the optimum case and thus provide significant cost savings under time-of-use pricing scenarios.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Measuring Neutron Polarisation in Deuteron Photo-disintegration with the CLAS Start Counter [Thesis]

Deuteron photo-disintegration (γd → γp) is a reaction that represents the simplest case in which nuclear and hadron physics models can be tested. Despite this, associated polarization analyses are limited in terms of angular coverage and energy ranges, especially in observables related to the recoil neutron. This is largely due to a lack in dedicated polarimetry equipment, and represents a roadblock in global progress to understand high-energy phenomena such as hexaquarks, and quark-gluon degrees of freedom. To address this problem, this PhD thesis pioneers a new methodology for the parasitic measurement of nucleon polarization using kinematic reconstruction of (spin-dependent) nucleon-nucleus scattering of reaction products, prior to their detection in large acceptance particle detector apparatus. Following this novel approach, which requires no dedicated polarimeter, a determination of the double polarization observable, $C^n_{x'}$, from deuteron photo-disintegration is presented, using Jefferson Lab’s CLAS detector. The analysis utilizes the (n,p) charge exchange reaction in CLAS’s "start counter" (plastic scintillator) to determine the final state neutron polarizations. The results present the first ever data for this observable above 0.7 GeV (photon beam energy) and significantly extend the angular range of the world data set. This new data is largely statistically consistent with the previous measurement of $C^n_{x'}$ by Bashkanov et al . in the overlapping energy range of 0.4-0.7 GeV. It is planned for the statistical accuracy of the presented result to be increased by the inclusion of additional data. The analysis herein serves as a key proof of concept for future applications, including a recommended similar analysis to be implemented with data from the more modern CLAS12 detector. This paves the way for a plethora of additional analyses using existing data sets that would provide crucial new constraints for hadron and nuclear physics.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

2D reactive transport model of shale chemical weathering and biogeochemical fluxes along a mountainous hillslope, East River Watershed, Colorado: Input files and simulation results

This data package contains input files and simulation results for a two-dimensional (2D) reactive transport model used to quantitatively analyze the coupled hydrological and biogeochemical processes governing shale weathering and associated biogeochemical fluxes under realistic environmental conditions in the high-elevation East River Watershed. These data support the conclusions presented in Stolze et al. (Water Resources Research, under review), "Model-based interpretation of solute exports and carbon partitioning during shale weathering in a mountainous hillslope". The model simulates atmospheric-subsurface gas exchange, subsurface water flow, and shale weathering processes under dynamic, year-scale conditions along a shale-underlain hillslope located in the East River watershed. The simulations were performed using the PFLOTRAN flow and reactive transport code and executed on the Perlmutter supercomputer to leverage its large-scale parallel computing capabilities. The data package contains two zipped folders, "model_input_files" and "simulation_results", and one readme.txt file. "model_input_files" contains the necessary input files to run the calibrated base-base model presented in Stolze et al. (Water Resources Research, under review). "simulation_results" contains a single hdf5 file ("Output_2D_hillslope_model.h5") which includes the results of simulation performed using the base-case model. This file can be opened with HDFView 3.1.4, Python, or MATLAB. "readme.txt" contains relevant information about the base-case model and provides guidelines on how to run the associated input files provided in the folder "model_input_files". Furthermore, readme.txt provides information regarding the model results provided in "Output_2D_hillslope_model.h5" such as matrix dimensionality and output units. Field datasets used to evaluate model performance were collected at three monitoring wells located along a hillslope transect (PLM1, PLM2, and PLM3). Dissolved ion concentration data were collected from November 2016 to October 2021 for Ca, Mg, DIC, Na, K, SO4 (Dong et al., 2025 - dic_npoc_data_2014_2024.zip - DOI:10.15485/1660459; Williams et al., 2025 - anion_data_2014_2024.zip - DOI:10.15485/1668054; Dong et al., 2025 - cation_data_2014_2024.zip - DOI:10.15485/1668055). Note that we used the files named er_PLM1_xx_yy, er_PLM2_xx_yy, and er_PLM3_xx_yy where xx stands for the name of the aqueous species and yy stands for the depth where the measurements were performed. Soil water content ([0 - 1] m) and water table depth were collected from November 2016 to October 2021 (Wan et al., 2024 - Dynamic_water_table__depthsFig2b.csv and Soil_water_content_Fig4e.csv - DOI:10.15485/2322567). Gaseous CO2 concentration were collected from October 2020 to December 2021(Wan et al., 2024 - Soil_CO2_concentrations_Fig4h.csv - DOI:10.15485/2322567) Gaseous CO2 flux from the subsurface to the atmosphere were collected in the vicinity of PLM2 from October 2019 to May 2022 (Wu et al., 2025). Soil microbial biomass concentration was measured from August 2016 to June 2017 (Sorensen et al., 2019 - 2017_East_River_Pumphouse_Microbial_Biomass__1_.csv - DOI:10.15485/1577267) All field data are published as CSV files compatible with Microsoft Excel, MATLAB, and Python, or as text files. The coordinates of the monitoring wells and the CO2(g) flux sensor in the coordinate system WGS84 are: -PLM1: [38.9197710 ; -106.9492750] -PLM2: [38.9201580 ; -106.9487170] -PLM3: [38.9207843 ; -106.9483668] -PLM4: 38.9210060 ; -106.9479528] -CO2(g) flux sensor: [38.9199180 ; -106.9489906] ------------------------------------------------------------------------------------------- This work was supported by the Watershed Function Science Focus Area at Lawrence Berkeley National Laboratory funded by the US Department of Energy, Office of Science, Biological and Environmental Research under Contract No. DE-AC02-05CH11231. This research used resources of the National Energy Research Scientific Computing Center (NERSC), a Department of Energy User Facility using NERSC award BER-ERCAP 23980, BER-ERCAP 28550, and BER-ERCAP 33789.

54 ENVIRONMENTAL SCIENCES↗

Greater aperture counteracts effects of reduced stomatal density on WUE: a case study on sugarcane and meta-analysis

Stomata regulate CO 2 and water vapor exchange between leaves and the atmosphere. Stomata are a target for engineering to improve crop intrinsic water use efficiency (iWUE). One example is by expressing genes that lower stomatal density (SD) and reduce stomatal conductance (g sw ). However, the quantitative relationship between reduced SD, g sw , and the mechanisms underlying it is poorly understood. We addressed this knowledge gap using low-SD sugarcane (Saccharum spp. hybrid) as a case study alongside a meta-analysis of data from 10 species. Transgenic expression of EPIDERMAL PATTERNING FACTOR 2 from Sorghum bicolor (SbEFP2) in sugarcane reduced SD by 26-38% but did not affect gsw compared to wildtype. Further, no changes occurred in stomatal complex size or proxies for photosynthetic capacity. Measurements of gas exchange at low CO 2 concentrations that promote complete stomatal opening to normalize aperture size between genotypes were combined with modeling of maximum gsw from anatomical data. These data suggest that increased stomatal aperture is the only possible explanation for maintaining gsw when SD is reduced. Meta-analysis across C 3 dicots, C 3 monocots, and C 4 monocots revealed engineered reductions in SD are strongly correlated with lower gsw (r 2 =0.60-0.98), but this response is damped relative to the change in anatomy.

59 BASIC BIOLOGICAL SCIENCES↗

NEPATEC2.0: NEPA Text Corpus v2.0

The National Environmental Policy Act of 1969, as amended (NEPA), is a major environmental law in the United States, requiring Federal agencies to consider and document potential environmental impacts before deciding on a proposed action. Modernization of NEPA and permitting processes faces significant challenges due to the lack of standardized formats and interoperable systems for organizing and sharing NEPA-related information across agencies. Much of the information gathered during NEPA reviews is written into documents such as categorical exclusions, environmental assessments, and environmental impact statements, then filed in predominately independent agency file stores that may or may not be publicly accessible. The application of metadata and data standards, such as those recommended by the Council on Environmental Quality (CEQ), to NEPA documents offers a shared vocabulary and structure for key entities like projects, processes, and documents that can streamline information exchange and enhance collaboration across systems. In this work, we publicly release NEPATEC2.0, an expanded corpus of NEPA documents with associated metadata. NEPATEC2.0 encompasses approximately 120,000 documents from 60,000 projects prepared by more than 60 different agencies. Modeled to align with CEQ metadata standards, NEPATEC2.0 promotes consistency in environmental reviews and supports the ongoing effort to modernize permitting technologies by facilitating more transparent, efficient, and data-driven decision-making. Importantly, NEPATEC2.0 demonstrates the possibilities and limitations of large language model-based prompting to extract information from NEPA documents at scale.

environmental review↗

NEPATEC v2.0: Standardized Metadata and Text Corpus of National Environmental Policy Act Documents

The National Environmental Policy Act of 1969, as amended (NEPA), is a major environmental law in the United States, requiring Federal agencies to consider and document potential environmental impacts before deciding on a proposed action. Modernization of NEPA and permitting processes faces significant challenges due to the lack of standardized formats and interoperable systems for organizing and sharing NEPA-related information across agencies. Much of the information gathered during NEPA reviews is written into documents such as categorical exclusions, environmental assessments, and environmental impact statements, then filed in predominately independent agency file stores that may or may not be publicly accessible. The application of metadata and data standards, such as those recommended by the Council on Environmental Quality (CEQ), to NEPA documents offers a shared vocabulary and structure for key entities like projects, processes, and documents that can streamline information exchange and enhance collaboration across systems. In this work, we publicly release NEPATEC2.0, an expanded corpus of NEPA documents with associated metadata. NEPATEC2.0 encompasses approximately 120,000 documents from 60,000 projects prepared by more than 60 different agencies. Modeled to align with CEQ metadata standards, NEPATEC2.0 promotes consistency in environmental reviews and supports the ongoing effort to modernize permitting technologies by facilitating more transparent, efficient, and data-driven decision-making. Importantly, NEPATEC2.0 demonstrates the possibilities and limitations of large language model-based prompting to extract information from NEPA documents at scale.

54 ENVIRONMENTAL SCIENCES↗

Representing cropping systems with the MEMS 2 ecosystem model

Abstract Croplands have been the focus of substantial investigation due to their considerable potential for sequestering carbon. Understanding the potential for soil organic carbon (SOC) sequestration and necessary management strategies will be enabled with accurate process‐based models. Accurately representing crop growth and agricultural practices will be critical for realistic SOC modeling. The MEMS 2 model incorporates a current understanding of SOC formation and stabilization, measurable SOC pools, and deep SOC dynamics and is seen as a highly promising tool to inform management intervention for SOC sequestration. Thus far, MEMS 2 has been developed to represent grasslands. In this study, we further developed MEMS 2 to model annual grain crops and common agricultural practices, such as irrigation, fertilization, harvesting, and tillage. Using four Ameriflux sites, we demonstrated an accurate simulation of crop growth and development. Model performance was strong for simulating aboveground biomass (index of agreement [d] range of 0.89–0.98) and green leaf area index (dfrom 0.90 to 0.96) across corn, soybean, and winter wheat. Good agreement with observations was also achieved for net ecosystem CO 2 exchange (dfrom 0.90 to 0.96), evapotranspiration (dfrom 0.91 to 0.94), and soil temperature (dof 0.96), while discrepancy with the available soil water content data remain (dfrom 0.14 to 0.81 at four depths to 100 cm). While we will continue model testing and improvement, MEMS 2 (version 2.14) has now demonstrated its ability to effectively simulate the growth of common grain crops and practices.

Agriculture↗

Explainable machine learning to quantify the value of proximal remote sensing in latent energy flux estimation

Proximal remote sensing has the potential to provide critical information on vegetation biophysical factors that can predict land-atmosphere exchange of water and energy. Latent energy (LE) flux is traditionally estimated using process-based models which rely on vegetation parameters that change during the growing season. Data-driven models have the potential to address these issues by offering flexible predictor selection and more efficient utilization of the information in predictor sets. These models require careful choice of predictors to avoid redundancy and allow robust cross-validation. In this study we present a systematic and comprehensive evaluation of machine learning (ML) models to assess the capability of meteorological and proximal sensing data for predicting LE at a half-hourly temporal resolution across multiple growing seasons for an agricultural system. The results presented here demonstrate that a model using four environmental predictors in combination with two proximal sensing variables can capture 88 % of the variability in LE. ML models using only three predictors (one meteorological and two proximal remote sensing) captured 81 % of LE variability, offering the best trade-off between performance and complexity. An ML model utilizing only two predictors, one proximal remote sensing variable and downwelling radiation, captured 77 % of LE variability. These results demonstrate the power of proximal remote sensing and meteorological observations to estimate land-atmosphere water vapor exchange, providing a solution where more direct methods such as eddy covariance are not available and for evaluations of agronomic management and genotypic variations.

60 APPLIED LIFE SCIENCES↗

Transfer Learning Meets Embedded Correlated Wavefunction Theory for Chemically Accurate Molecular Simulations: Application to Calcium Carbonate Ion Pairing

Achieving chemical accuracy for molecular simulations remains a central challenge in computational chemistry. Here, we present an embedded correlated wavefunction transfer learning (ECW-TL) framework for accurately simulating molecular dynamics in the condensed phase. ECW-TL incorporates high-level electron exchange and correlation effects in ECW theory while preserving the training and computational efficiency of machine-learned interatomic potentials. We demonstrate the framework on Ca 2+ –CO 3 2– ion pairing in aqueous solution, a key process underlying CO 2 mineralization in seawater. As proof of principle, we first show that fine-tuning a DFT-revPBE-D3(BJ) baseline model with embedded-DFT-SCAN data reproduces the DFT-SCAN free-energy surface within 1 kcal/mol across all solvation states. Extending the framework to embedded MP2 and localized natural-orbital CCSD(T) further refines the free-energy profile, revealing the crucial role of exact electron exchange and correlation in determining ion-pair stability and structure. The computed ion-pair association free energy is in quantitative agreement with experimental measurements, further validating the accuracy of the ECW-TL framework. ECW-TL thus provides a general, data-efficient route for transferring CW accuracy to efficient simulations of complex aqueous and interfacial chemical processes.

cluster chemistry↗