Search NASASearch

SEARCH · Search NASA

Results for “model set up”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Evaluation of dynamical downscaling in a fully coupled regional earth system model

A set of decadal simulations has been completed and evaluated for gains using the Regional Arctic System Model (RASM) to dynamically downscale data from a global Earth system model and two atmospheric reanalyses. RASM is a fully coupled atmosphere–land–ocean–sea ice regional Earth system model. Nudging to the forcing data is applied to approximately the top half of the atmospheric domain. RASM simulations were also completed with a modification to the atmospheric physics for evaluating changes to the modeling system. The results show that for the top half of the atmosphere, the RASM simulations follow closely to that of the forcing data, regardless of the forcing data. The results for the lower half of the atmosphere, as well as the surface, show a clustering of atmospheric state and surface fluxes based on the modeling system. At all levels of the atmosphere the imprint of the weather from the forcing data is present as indicated in the pattern of the annual means. Biases, in comparison to reanalyses, are evident in the Earth system model forced simulations for the top half of the atmosphere but are not present in the lower atmosphere. This suggests that bias correction is not needed for fully coupled dynamical downscaling simulations. While the RASM simulations tended to go to the same mean state for the lower atmosphere, there are a differences in the variability and changes of weather patterns across the ensemble of simulations. These differences in the weather result in variances in the sea ice and oceanic states.

54 ENVIRONMENTAL SCIENCES

Cooperative Agreement To Analyze variabiLity, change and predictabilitY in the earth SysTem (CATALYST)

CATALYST proposes to perform foundational coordinated research in a team-oriented collaborative effort aimed at advancing a robust understanding of modes of Earth system variability and change using models, observations and process studies. The proposed research will address the DOE/BER mission by exploring the limits to predictability, identifying fundamental underlying mechanisms, quantifying interactions among modes of variability, and discovering tipping points in the Earth system to understand the current and future impacts of these phenomena on regional and global climate. Four fundamental gaps are identified in our knowledge of the Earth system: 1) What are the limits to predictability on various timescales? 2) What are the interactions among modes of Earth system variability? 3) How may modes of Earth system variability change in response to changes in external forcing, and what are the tipping points involved with those changes? 4) How are high impact events connected to modes of Earth system variability and how may they change in the future? Related to those gaps in our knowledge, we formulate four research objectives to address those gaps using a combination of Earth system models (ESMs) and machine learning (ML) methods. Research Objective 1 (RO1) addresses the first gap above and proposes to understand modes of variability and their limits of predictability on subseasonal to decadal timescales using ESMs and ML. Research Objective 2 (RO2) addresses the second gap and proposes to use a hierarchy of models to understand relevant processes and feedbacks related to how modes of variability interact with each other. Research Objective 3 (RO3) is designed to study the third gap and proposes to examine the role of external forcings in changes of modes of Earth system variability and their interactions, and the likelihood and predictability of tipping points and irreversible changes. Research Objective 4 (RO4) will address the fourth gap and proposes to use high resolution ESMs, regionally refined models (RRMs), and ML methods to investigate the relationships between high impact events (e.g. flash droughts and precipitation extremes, atmospheric rivers (ARs), tropical cyclones (TCs), storm surge/sea level rise), the synoptic systems that produce them, and their changes related to modes of Earth system variability. The research will involve the use of the Community Earth System Model (CESM), Energy Exascale Earth System Model (E3SM), CMIP multi-model data sets, a hierarchy of simpler models, and numerous observational data sets. In the course of the proposed research, CATALYST will contribute to metrics and diagnostics that will be integrated in Coordinated Model Evaluation Capabilities (CMEC), particularly with regards to the Quasi-biennial Oscillation (QBO) and its interactions with the Madden-Julian Oscillation (MJO), high atmospheric pressure blocking, and new precipitation metrics.

54 ENVIRONMENTAL SCIENCES

An Entropy-Based Test and Development Framework for Uncertainty Modeling in Level-Set Visualizations

We present a simple comparative framework for testing and developing uncertainty modeling in uncertain marching cubes implementations. The selection of a model to represent the probability distribution of uncertain values directly influences the memory use, run time, and accuracy of an uncertainty visualization algorithm. We use an entropy calculation directly on ensemble data to establish an expected result and then compare the entropy from various probability models, including uniform, Gaussian, histogram, and quantile models. Our results verify that models matching the distribution of the ensemble indeed match the entropy. We further show that fewer bins in nonparametric histogram models are more effective whereas large numbers of bins in quantile models approach data accuracy.

Sisneros, Robert

Exploring the Whole Set of Accurate Sparse Interpretable Models

In data science applications, there are often many models that fit the data well. This phenomenon was called the Rashomon Effect by Leo Breiman. The set of good models is called the Rashomon Set, and the goal of this project is to locate, store, and study the Rashomon sets for classes of interpretable models, including decision trees and generalized additive models.

97 MATHEMATICS AND COMPUTING

Assessment of the CTF subchannel code for modeling a large-break loss-of-coolant accident reflood transient

With increased industry interest in extending reactor operating cycles, the Nuclear Energy Advanced Modeling and Simulation (NEAMS) program has been investigating the behavior of high-burnup fuel during design basis accidents such as the large-break loss-of-coolant accident (LBLOCA) with consideration for risk of fuel fragmentation, relocation, and dispersal (FFRD). As part of that activity, the NEAMS subchannel thermal/ hydraulics (T/H) code, CTF, is being used for modeling of LBLOCA and to determine the impact of subchannel resolution on results. Although CTF includes a wide range of models for LBLOCA conditions, the code has not been used for this application while maintained at Oak Ridge National Laboratory (ORNL) until now. Therefore, here, in this work, a preliminary assessment of several of these models was performed using openly available reflood experimental data from the Flooding Experiments in Blocked Arrays (FEBA) tests. One coarse mesh and one fine mesh model were set up in CTF for high and low flooding rate tests performed in the unblocked FEBA facility. A coarse TRACE model was set up to be as consistent as possible with the coarse CTF model to allow for code-to-code benchmarking. The assessment shows a tendency of the codes to over-predict peak cladding temperature (PCT) near the top of the bundle and to quench early. Advanced spacer grid models were shown to improve upper bundle predictions in CTF. The resolved CTF model over-predicted PCT by a larger degree in the center channels in the low-flooding rate test, and it is believed that the radiative heat transfer model, which was not used in this study, may be needed to correct this over-prediction. Finally, this work demonstrates the importance of the droplet model in determining quench time and vapor temperature and PCT prediction, which necessitates a more in-depth validation of these models in the future.

21 SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLANTS

An adaptive model-free robotic force control strategy for hydrodynamic real-time hybrid simulation of floating offshore wind turbines

Real-time hybrid simulation (RTHS) - a cyber-physical testing approach - promises to enhance the simulation fidelity of the model-scale experiments used to prototype floating offshore wind turbines (FOWTs). In hydrodynamic RTHS (hydro-RTHS), actuators emulate aerodynamic forces on model-scale FOWT specimens subjected to physical waves in a hydrodynamic laboratory. Robotic arms are promising candidates for actuation in hydro-RTHS due to their compact multi-degree-of-freedom (DOF) capabilities. Unlike classical RTHS for seismic applications, which typically relies on displacement control, hydro-RTHS requires 6-DOF force control on newly designed floating prototypes in a model-scale setting, which presents significant challenges, including modeling uncertainties, directional asymmetry, configuration drift, bandwidth limitations, and time-varying delays. To mitigate these constraints without extensive pre-test calibration, this study proposes an adaptive model-free robotic force control strategy that combines task-space explicit force control with a secondary joint-space pose-keeping task. The Adaptive Feedforward Compensator (AFC) is integrated into the force control loop to compensate for time-varying delay. Experimental testing was conducted using a Franka Emika Panda robotic arm with a 1:50 scale FOWT specimen under operational wind and wave conditions. Results demonstrate stable and consistent 6-DOF force tracking. Effective delay compensation was observed, with low-frequency delay reductions ranging from 71.4% to 91.8% and improvements in low-frequency surge force tracking of 25.0% to 52.1%. This study enhances robotic actuation performance in hydro-RTHS and introduces a force control strategy that supports reliable robotic operation in uncertain floating environments. Future work will explore disturbance-observer mechanisms to further enhance wave rejection capabilities under extreme wind and wave conditions.

17 WIND ENERGY

AI-Batt (Autonomous Identification of Battery Life Models) [SWR 21-36]

Autonomous Identification of Battery Life Models (AI-Batt) AI-Batt is a MATLAB code base for developing lifetime models for batteries from accelerated aging data. The code base provides many functions for processing, visualizing, and modeling battery aging data, making the data processing, exploration, and modeling workflow substantially faster. These tools are tailored for working with battery aging data sets, which usually consist of many separate time-series for each cell, with many test conditions and possible replicates at each condition, which makes it difficult to simply process or visualize the data set. Complex modeling tasks, such as cross-validation, sensitivity analysis, and uncertainty quantification have been implemented to enable thorough statistical investigation of model predictions. Additionally, several machine-learning algorithms are implemented to autonomously identify suitable models via symbolic regression. Data processing functions automatically cast data from the struct data type, which is commonly used to store experimental data, but is not an acceptable input for most algorithms, to the table data type, which can be easily used as input to any optimization algorithm. Also, the data can be separated into time-invariant and time-variant data tables, which is helpful for exploring the data set as well as developing separate models for time-variant and time-invariant aging mechanisms. For example, in aging tests with constant temperature, temperature is a time-invariant experimental condition. Visualization tools enable plotting of data, model fits, and model simulations possible with single-line function calls, empowering data exploration of complex data sets with both time-varying and time-invariant trends. Plots can be automatically generated for the whole data set, or separated by data group (groups of test replicates) or individual data series. Data points or data series can be automatically colored by the value of a variable with a variety of color maps, and model predictions can also be colored by the value of a fit statistic. Comparisons between data sets and the predictions/simulations of different models on the same data set can be easily plotted as well. Distributions of parameter values from bootstrap resampling can be plotted to visualize the reliability of parameter estimation, or determine any correlations between parameters. Modeling tools handle the complex task of creating and parsing symbolic equations for modeling battery lifetime. Equations are parsed to grab relevant data variables, parameter values, or specified sub-models for input into optimization, evaluation, or simulation functions. Models can be optimized locally (one set of parameters for each data series), bi-level (some parameters shared across the data set), or globally (single set of parameters for all data). Functions implementing symbolic regression algorithms help users to discover effective model equations, even in poorly sampled, high-dimensional data.

Smith, Kandler [National Renewable Energy Lab. (NR

Search for displaced leptons in $\sqrt{𝑠}$ =13 TeV and 13.6 TeV 𝑝⁢𝑝 collisions with the ATLAS detector

A search for leptons displaced from the primary vertex is performed with the ATLAS detector at the Large Hadron Collider. The search includes the full proton-proton collision dataset collected during Run 2 at $\sqrt{𝑠}$ = 13 TeV and a partial dataset collected during Run 3 in 2022–2023 at $\sqrt{𝑠}$ = 13.6 TeV, corresponding to integrated luminosities of 140 fb −1 and 56.3 fb −1 , respectively. Final states with displaced electrons or muons are considered, and novel triggers introduced in Run 3 are employed that use large impact parameter tracking to reconstruct displaced tracks with low momentum. In addition, photon reconstruction and multivariate techniques are employed to broaden the sensitivity to channels with large background rates or highly displaced electrons, respectively. The results are consistent with the Standard Model background expectations and are used to set model-independent limits on the production of displaced electrons and muons. The analysis is also interpreted in the context of a gauge-mediated supersymmetry breaking model with pair-produced long-lived sleptons and a dark sector model with pair-produced chargino-like states. The results include 95% confidence level exclusions of selectrons with lifetimes from 4 ps to 60 ns and a mass of 150 GeV, and exclusions of selectrons, smuons, and staus with a lifetime of 0.3 ns for masses up to 740, 830, and 440 GeV, respectively. Dark charginos with masses up to 380 GeV are excluded for a mass difference with the neutral state of 40 GeV, and mass differences down to 17 GeV are excluded for dark charginos with a 100 GeV mass.

supersymmetric models

Bayesian model mixing with multireference energy density functional

Reliably predicting nuclear properties across the entire chart of isotopes is important for applications ranging from nuclear astrophysics to superheavy science to nuclear technology. To this day, however, all the theoretical models that can scale at the level of the chart of isotopes remain semiphenomenological. Because they are fitted locally, their predictive power can vary significantly; different versions of the same theory provide different predictions. Bayesian model mixing takes advantage of such imperfect models to build a local mixture of a set of models to make improved predictions. Earlier attempts to use Bayesian model mixing for mass table calculations relied on models treated at single-reference energy density functional level, which fail to capture some of the correlations caused by configuration mixing or the restoration of broken symmetries. In this study we have applied Bayesian model mixing techniques within a multireference energy density functional (MR-EDF) framework. We considered predictions of two-particle separation energies from particle number projection or angular momentum projection with four different energy density functionals—a total of eight different MR-EDF models. We used a hierarchical Bayesian stacking framework with a Dirichlet prior distribution over weights together with an inverse log-ratio transform to enable positive correlations between different models. We found that Bayesian model mixing provides significantly improved predictions compared to the participating models. Published by the American Physical Society 2025

73 NUCLEAR PHYSICS AND RADIATION PHYSICS

ARM Data-Oriented Metrics and Diagnostics Package for Climate Model Evaluation

A Python-based metrics and diagnostics package is currently being developed by the U.S. Department of Energy (DOE) Atmospheric Radiation Measurement (ARM) Infrastructure Team at Lawrence Livermore National Laboratory (LLNL) to facilitate the use of long-term, high-frequency measurements from the ARM Facility in evaluating the regional climate simulation of clouds, radiation, and precipitation. This metrics and diagnostics package computes climatological means of targeted climate model simulation and generates tables and plots for comparing the model simulation with ARM observational data. The Coupled Model Intercomparison Project (CMIP) model data sets are also included in the package to enable model intercomparison as demonstrated in Zhang et al. (2017). The mean of the CMIP model can serve as a reference for individual models. Basic performance metrics are computed to measure the accuracy of mean state and variability of climate models. The evaluated physical quantities include cloud fraction, temperature, relative humidity, cloud liquid water path, total column water vapor, precipitation, sensible and latent heat fluxes, and radiative fluxes, with plan to extend to more fields, such as aerosol and microphysics properties. Process-oriented diagnostics focusing on individual cloud- and precipitation-related phenomena are also being developed for the evaluation and development of specific model physical parameterizations. The version 1.0 package is designed based on data collected at ARM’s Southern Great Plains (SGP) Research Facility, with the plan to extend to other ARM sites. The metrics and diagnostics package is currently built upon standard Python libraries and additional Python packages developed by DOE (such as CDMS and CDAT). The ARM metrics and diagnostic package is available publicly with the hope that it can serve as an easy entry point for climate modelers to compare their models with ARM data. In this report, we first present the input data, which constitutes the core content of the metrics and diagnostics package in section 2, and a user's guide documenting the workflow/structure of the version 1.0 codes, and including step-by-step instruction for running the package in section 3.

54 ENVIRONMENTAL SCIENCES

Impact of T - and ρ -dependent decay rates and new (n, γ ) cross-sections on the s process in low-mass asymptotic giant branch stars

Aims. We study the impact of nuclear input related to weak-decay rates and neutron-capture reactions on predictions for the slow neutron-capture process (s process) in asymptotic giant branch (AGB) stars. We provide the first database of surface abundances and stellar yields of the isotopes heavier than iron from the Monash models. Methods. We ran nucleosynthesis calculations with the Monash post-processing code for seven stellar structure evolution models of low-mass AGB stars with three different sets of nuclear inputs. The reference set has constant decay rates and represents the set used in the previous Monash publications. The second set contains the temperature and density dependence of β decays and electron captures based on the default rates of nuclear NETwork GENerator (NETGEN). In the third set, we further update 92 neutron-capture rates based on re-evaluated experimental cross sections from the ASTrophysical Rate and rAw data Library. We compare and discuss the predictions of the sets relative to each other in terms of isotopic surface abundances and total stellar yields. We also compare the results to isotopic ratios measured in presolar stardust silicon carbide (SiC) grains from AGB stars. Results. The new sets of models result in a ∼66% solar s-process contribution to the p-nucleus 152 Gd, confirming that this isotope is predominantly made by the s process. The nuclear input updates result in predictions for the 80 Kr/ 82 Kr ratio in the He intershell and surface 64 Ni/ 58 Ni, 94 Mo/ 96 Mo, and 137 Ba/ 136 Ba ratios that are more consistent with the corresponding ratios measured in stardust; however, the new predicted 138 Ba/ 136 Ba ratios are higher than the typical values of the SiC grains. The W isotopic anomalies are in agreement with data from the analyses of other meteoritic inclusions. We confirm that the production of 176 Lu and 205 Pb is affected by too large uncertainties in their decay rates from NETGEN.

79 ASTRONOMY AND ASTROPHYSICS

Challenges and Opportunities for Electric Utility Modeling and Asset Valuation Frameworks: Case Study on Valuing New Pumped Storage Hydropower

Asset valuation by electric utilities is becoming increasingly difficult in the rapidly changing electric sector. Rapid deployment of variable generation and inverter-based storage systems along with uncertain demand growth, climate, policies, and other factors create a challenging environment for understanding the value proposition of a new potential asset. This report describes an effort between the Tennessee Valley Authority (TVA) and three U.S. Department of Energy laboratories to perform a detailed review of utility modeling and analysis practices for asset valuation and identify challenges and opportunities for advancing its methods into the future. It focuses on a case study of new potential pumped storage hydropower (PSH) because of growing interest in new PSH capacity to provide energy balancing, firm capacity, and a range of ancillary services. Staff from the DOE labs conducted systematic interviews about current practices in capacity expansion modeling, production-cost modeling, hydrological modeling, and transmission stability modeling while also discussing how scenario analysis is conducted and how models and data are integrated. The effort resulted in a set of model, integration, and scenario recommendations that could be valuable to TVA, other utilities, system operators, and other stakeholders conducting integrated grid analysis. Individual model recommendations suggest exploring computational tradeoffs with detail and resolution across spatiotemporal structure, supply- and demand-side details, transmission overlays, market interactions, and ancillary services. Automated processes to pass data between models and conduct larger scenario suites could also enhance valuation practices by enabling a more consistent study of asset value across a broader range of uncertain future grid conditions where PSH could be particularly valuable. TVA and other industry stakeholders can learn from and adapt applied research-grade methods developed by DOE laboratories and other research institutions to improve decision making and accelerate progress towards a reliable, economic, sustainable energy system.

13 HYDRO ENERGY

Route Energy Prediction (RouteE) Powertrain Validation Report

The National Renewable Energy Laboratory's flagship package in the RouteE suite, RouteE-Powertrain, is a mesoscopic energy model that predicts vehicle energy consumption given discrete attributes that describe each segment or link in a vehicle's path on a road network. High-frequency, physics-based, powertrain simulators, such as NREL's FASTSim, are well-suited to model vehicle energy consumption when real driving data and a detailed understanding of the vehicle powertrain specifications are available. However, there are a variety of situations in the past, present (real-time), and future where high-frequency driving data and/or vehicle information may not be available, but reliable energy consumption is still desired, such as energy-aware vehicle routing. These are the ideal applications for RouteE-Powertrain. The suite of RouteE tools also includes RouteE-Compass, which is an eco-routing software that incorporates energy consumption into network routing algorithms, and RouteE-Mobile, which is a prototype smartphone navigation app to demonstrate the integrated capabilities of the RouteE suite for real-world eco-routing. The focus of this validation report is to share key metrics about the data sets and models behind RouteE-Powertrain. The set of RouteE-Powertrain models discussed in this report are made available through the RouteE web API through the NREL Developer Network.

33 ADVANCED PROPULSION SYSTEMS

Efficient First-Order Algorithms for Large-Scale, Non-Smooth Maximum Entropy Models with Application to Wildfire Science

Maximum entropy (MaxEnt) models are a class of statistical models that use the maximum entropy principle to estimate probability distributions from data. Due to the size of modern data sets, MaxEnt models need efficient optimization algorithms to scale well for big data applications. State-of-the-art algorithms for MaxEnt models, however, were not originally designed to handle big data sets; these algorithms either rely on technical devices that may yield unreliable numerical results, scale poorly, or require smoothness assumptions that many practical MaxEnt models lack. In this paper, we present novel optimization algorithms that overcome the shortcomings of state-of-the-art algorithms for training large-scale, non-smooth MaxEnt models. Our proposed first-order algorithms leverage the Kullback–Leibler divergence to train large-scale and non-smooth MaxEnt models efficiently. For MaxEnt models with discrete probability distribution of n elements built from samples, each containing m features, the stepsize parameter estimation and iterations in our algorithms scale on the order of O(mn) operations and can be trivially parallelized. Moreover, the strong ℓ1 convexity of the Kullback–Leibler divergence allows for larger stepsize parameters, thereby speeding up the convergence rate of our algorithms. To illustrate the efficiency of our novel algorithms, we consider the problem of estimating probabilities of fire occurrences as a function of ecological features in the Western US MTBS-Interagency wildfire data set. Our numerical results show that our algorithms outperform the state of the art by one order of magnitude and yield results that agree with physical models of wildfire occurrence and previous statistical analyses of wildfire drivers.

Physics

Power Profile Monitoring and Tracking Evolution of System-Wide HPC Workloads

The power & energy demands of HPC machines have grown significantly. Modern exascale HPC systems require tens of megawatts of combined power for computing resources and cooling facilities at full capacity. The current energy trend is not sustainable for future HPC systems, and there is a need to work toward the energy efficiency aspect of HPC performance. Energy awareness of the HPC applications at the job level is essential for running an efficient HPC system. This work aims to develop a pipeline to provide a production-level system-wide overview of the HPC workloads' power profile while handling evolving workloads exhibiting new power trends. We developed an open-set classification model for HPC jobs based on the properties of power profiles to continuously provide a system-wide holistic view of recently completed jobs. The pipeline helps continuously monitor the job-level power usage pattern of HPC and enables us to capture the new trends in applications' power behavior. We employed a comprehensive set of techniques to generate job-level data, custom-designed feature extraction methods to extract critical features from jobs' power profiles, clustering techniques powered by generative modeling, and open-set classification for identifying job profiles into known classes or an unknown set. With extensive evaluations, we demonstrate the effectiveness of each component in our pipeline. We provide an analysis of the resulting clusters that characterize the power profile landscape of the Summit supercomputer from more than 60K jobs executed in a year. The open-set classification classifies the known data sets into known classes with high accuracy and identifies unknown data noints with over 85% accuracy.

Karimi, Ahmad Maroof

The National Climate Data Base (NCDB): A Bias-Corrected High-Resolution Climate Dataset

Assessing renewable energy resources under future climate scenarios has been highlighted in recent years to analyze and understand potential impacts of future change in renewable generation on the power sector. Solar energy is well-known as the most plentiful among various renewable resources and usually converted to electricity using photovoltaics (PV) technologies, and the global deployment of PV technology has increased rapidly in recent decades. In this study, we develop a statistical technique to downscale the future projection of solar irradiance for PV energy-related applications. A set of Regional Climate Model (RCM)-based projections obtained from the North American Coordinated Regional Climate Downscaling Experiment (NA-CORDEX) are used as inputs to statistical methods to generate high-resolution global horizontal irradiance (GHI) over the contiguous United States (CONUS). The main steps of the statistical downscaling method include (1) regridding RCM output (0.22 degree and daily resolutions) to handle the modeled-observed data sets on a common grid, (2) correcting bias of RCM GHI using satellite-derived observation, and (3) implementing temporal and spatial downscaling to generate GHI at 8-km and hourly resolution. Basically, complex physical processes and interactions between solar radiation and various atmospheric constituents lead solar irradiance to be highly variable and uncertain. Underrepresentation of clouds from the RCM parameterizations is the main source of error and uncertainty in modeling solar irradiance. Thus, we adapt and use the high-quality satellite-derived data from the National Solar Radiation Database (NSRDB) to analyze the bias and error of RCM GHI as well as estimate the statistical parameters for spatial and temporal downscaling. This presentation will summarize the comprehensive analysis conducted to produce and assess the results under two climate scenarios (RCP4.5 and RCP8.5). We will also present a detailed validation demonstrating the strengths of the proposed downscaling method and future extension of this research.

climate data

Non-linear relationships between daily temperature extremes and US agricultural yields uncovered by global gridded meteorological datasets

Global agricultural commodity markets are highly integrated among major producers. Prices are driven by aggregate supply rather than what happens in individual countries in isolation. Furthermore, estimating the effects of weather-induced shocks on production, trade patterns and prices hence requires a globally representative weather data set. Recently, two data sets that provide daily or hourly records, GMFD and ERA5-Land, became available. Starting with the US, a data rich region, we formally test whether these global data sets are as good as more fine-scaled country-specific data in explaining yields and whether they estimate similar response functions. While GMFD and ERA5-Land have lower predictive skill for US corn and soybeans yields than the fine-scaled PRISM data, they still correctly uncover the underlying non-linear temperature relationship. All specifications using daily temperature extremes under any of the weather data sets outperform models that use a quadratic in average temperature. Correctly capturing the effect of daily extremes has a larger effect than the choice of weather data. In a second step, focusing on Sub Saharan Africa, a data sparse region, we confirm that GMFD and ERA5-Land have superior predictive power to CRU, a global weather data set previously employed for modeling climate effects in the region.

54 ENVIRONMENTAL SCIENCES