Search NASA⌕ Search

SEARCH · Search NASA

Results for “Statistical error”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

250 records · Page 14

Stochastic Optimization and Uncertainty Quantification of Natrium-based Nuclear-Renewable Energy Systems for Flexible Power Applications in Deregulated Markets

Rapid integration of variable renewable energy sources (VRES) has made modeling and stochastic optimization of hybrid energy systems crucial for studying their long-term performance and viability. However, most studies have focused on just historical data, which may be unreliable for capturing short-term fluctuations, rare events, and long-term patterns of energy demand, price, and the variability of renewable energy sources. For this study, optimal synthetic time series models were developed using Wasserstein distance. The models were validated by comparing the key statistical measures against those of the historical data. They were then used to optimize the integrated Natrium-style advanced energy systems and their long-term (30 years) economics. The stochastic model performs bi-level optimization to find the optimal sizes for the balance of plant and thermal energy storage, while also optimizing energy dispatch to achieve the maximum net present value. In studies of two deregulated markets (California ISO and the Electric Reliability Council of Texas), the integrated Natrium-style system performed better in CAISO than in ERCOT, given higher and more consistent electricity prices during peak-demand periods. The potentially enlarged cost associated with the variable operation and maintenance of the TES system also plays a significant role in driving the system sizing, thus its impacts on the system are investigated in detail through comparison against a baseline case. The study also finds that the bi-level optimization results based on stochastic gradient descent closely match the grid search results. The uncertainty quantification of the stochastic signals provides further NPV-related insights and probability distributions for the case studies. The normal standard error of the mean of NPV for the case with and without TES VOM for CAISO were found to be 7.73M (plus-minus sign) 1.09M USD and 104.99M (plus-minus sign) 1.25M USD, respectively based on a 95% confidence. Given the relatively small NPV variance based on 150 samples, the analysis affords the most robust possible prediction of the techno-economic performance of the integrated Natrium-style energy systems.

25 ENERGY STORAGE↗

Prediction of hydration energies of adsorbates at Pt(111) and liquid water interfaces using machine learning

Aqueous phase heterogeneous catalysis is important to various industrial processes, including biomass conversion, Fischer–Tropsch synthesis, and electrocatalysis. Accurate calculation of solvation thermodynamic properties is essential for modeling the performance of catalysts for these processes. Explicit solvation methods employing multiscale modeling, e.g., involving density functional theory and molecular dynamics have emerged for this purpose. Although accurate, these methods are computationally intensive. This study introduces machine learning (ML) models to predict solvation thermodynamics for adsorbates on a Pt(111) surface, aiming to enhance computational efficiency without compromising accuracy. In particular, ML models are developed using a combination of molecular descriptors and fingerprints and trained on previously published water–adsorbate interaction energies, energies of solvation, and free energies of solvation of adsorbates bound to Pt(111). These models achieve root mean square error values of 0.09 eV for interaction energies, 0.04 eV for energies of solvation, and 0.06 eV for free energies of solvation, demonstrating accuracy within the standard error of multiscale modeling. Feature importance analysis reveals that hydrogen bonding, van der Waals interactions, and solvent density, together with the properties of the adsorbate, are critical factors influencing solvation thermodynamics. Furthermore, these findings suggest that ML models can provide rapid and reliable predictions of solvation properties. This approach not only reduces computational costs but also offers insights into the solvation characteristics of adsorbates at Pt(111)–water interfaces.

Adsorption↗

Role of the likelihood for elastic scattering uncertainty quantification

In the last decade, uncertainty quantification (UQ) for optical model potentials (OMPs) has become a focal point for nuclear reaction theory, and several competing approaches for OMP UQ have recently been developed. Here, we clarify recent efforts to compare frequentist and Bayesian approaches in the context of OMP UQ [G. B. King et al., Phys. Rev. Lett. 122, 232502 (2019)]. We replicate a portion of that OMP UQ study but use independent statistical tools. Specifically, we compare two methods for OMP parameter inference from elastic scattering data: the Levenberg-Marquardt algorithm for χ 2 minimization on one hand and Markov chain Monte Carlo (MCMC) sampling on the other. Separately, we assess the common practice of using a renormalized likelihood (χ 2 /N), N being the number of data points, instead of the canonical weighted-least-squares likelihood (χ 2 ), as a way of accounting for unknown data correlations. Here, we show that for a generic linear model and for a five-parameter OMP analysis, frequentist and uniform-prior Bayesian approaches recover the same optimum and uncertainty estimates—not systematically larger uncertainties for the Bayesian approach, as was concluded in G. B. King et al., Phys. Rev. Lett. 122, 232502 (2019). Further, we show that if an additional, near-degenerate parameter is introduced into the same OMP analysis such that the parameter posterior becomes non-Gaussian, then covariance-based estimates of uncertainty become unreliable. Finally, we show that regardless of optimization approach, if χ 2 /N is used for the likelihood, the resulting parametric uncertainties increase by $\sqrt{N}$, and that this is responsible for the conclusions drawn in the revisited study. Based on our replication results, we find that a fortuitous cancellation of unreported errors and the renormalization factor can lead to improvement in empirical coverages, as was the case in the original comparative study. We emphasize that developing and applying a realistic likelihood function is an essential task in a UQ analysis, and that several recent UQ studies that employed a renormalized likelihood (i.e., including a 1/N factor) may have yielded unrealistically large uncertainties for elastic-scattering observables. If the parameter posterior deviates from multivariate-normal, a sampling-based approach like MCMC has a clear advantage over methods that assume the Laplace approximation holds. We note that empirical coverage can serve as an important internal check for the analyst whose model or data may have additional, unaccounted-for uncertainties.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

A hybrid neural architecture: Online attosecond x-ray characterization

The emergence of high-repetition-rate x-ray free-electron lasers (XFELs), such as SLAC’s LCLS-II, serves as our canonical example for autonomous controls that necessitate high-throughput diagnostics paired with streaming computational pipelines capable of single-shot analysis with extremely low latency. We present the deterministic characterization with an integrated parallelizable hybrid resolver architecture, a hybrid machine learning framework designed for fast, accurate analysis of XFEL diagnostics using angular streaking-based sinogram images. This architecture integrates convolutional neural networks and bidirectional long short-term memory models to denoise input, identify x-ray sub-spike features, and extract sub-spike relative delays with sub-30 attosecond temporal resolution. Deployed on low-latency hardware, it achieves over 10 kHz throughput with 168.3 μs inference latency, indicating scalability to 14 kHz with field-programmable gate array integration. By transforming regression tasks into classification problems and leveraging optimized error encoding, we achieve high precision with low-latency performance that is critical for real-time streaming event selection and experimental control feedback signals. This represents a key development in real-time control pipelines for next-generation autonomous science, generally, and high repetition-rate x-ray experiments in particular.

Accelerator Physics (physics.acc-ph)↗

Navigating the Noise: Bringing Clarity to ML Parameterization Design With O $\boldsymbol{\mathcal{O}}$(100) Ensembles

Abstract Machine‐learning (ML) parameterizations of subgrid processes (here of turbulence, convection, and radiation) may one day replace conventional parameterizations by emulating high‐resolution physics without the cost of explicit simulation. However, uncertainty about the relationship between offline and online performance (i.e., when integrated with a large‐scale general circulation model) hinders their development. Much of this uncertainty stems from limited sampling of the noisy, emergent effects of upstream ML design decisions on downstream online hybrid simulation. Our work rectifies the sampling issue via the construction of a semi‐automated, end‐to‐end pipeline for size ensembles of hybrid simulations, revealing important nuances in how systematic reductions in offline error manifest in changes to online error and online stability. For example, removing dropout and switching from a Mean Squared Error to a Mean Absolute Error loss both reduce offline error, but they have opposite effects on online error and online stability. Other design decisions, like incorporating memory, converting moisture input from specific humidity to relative humidity, using batch normalization, and training on multiple climates do not come with any such compromises. Finally, we show that ensemble sizes of may be necessary to reliably detect causally relevant differences online. By enabling rapid online experimentation at scale, we can empirically settle debates regarding subgrid ML parameterization design that would have otherwise remained unresolved in the noise.

Lin, Jerry [Department of Earth System Sciences Un↗

X-ray imaging and electron temperature evolution in laser-driven magnetic reconnection experiments at the national ignition facility

We present results from x-ray imaging of high-aspect-ratio magnetic reconnection experiments driven at the National Ignition Facility. Two parallel, self-magnetized, elongated laser-driven plumes are produced by tiling 40 laser beams. A magnetic reconnection layer is formed by the collision of the plumes. A gated x-ray framing pinhole camera with micro-channel plate detector produces multiple images through various filters of the formation and evolution of both the plumes and current sheet. As the diagnostic integrates plasma self-emission along the line of sight, two-dimensional electron temperature maps ⟨Te⟩Y are constructed by taking the ratio of intensity of these images obtained with different filters. The plumes have a characteristic temperature ⟨Te⟩Y=240 ± 20 eV at 2 ns after the initial laser irradiation and exhibit a slow cooling up to 4 ns. The reconnection layer forms at 3 ns with a temperature ⟨Te⟩Y=280 ± 50 eV as the result of the collision of the plumes. The error bars of the plumes and current sheet temperatures separate at 4 ns, showing the heating of the current sheet from colder inflows. Using a semi-analytical model, we survey various heating mechanisms in the current sheet. We find that reconnection energy conversion would dominate at low density (ne≲7×1018 cm−3) and electron-ion collisional drag at high-density (≳1019 cm−3).

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Recovered supernova Ia rate from simulated LSST images

Aims.TheVera C. RubinObservatory’s Legacy Survey of Space and Time (LSST) will revolutionize time-domain astronomy by detecting millions of different transients. In particular, it is expected to increase the number of known type Ia supernovae (SN Ia) by a factor of 100 compared to existing samples up to redshift ∼1.2. Such a high number of events will dramatically reduce statistical uncertainties in the analysis of the properties and rates of these objects. However, the impact of all other sources of uncertainty on the measurement of the SN Ia rate must still be evaluated. The comprehension and reduction of such uncertainties will be fundamental both for cosmology and stellar evolution studies, as measuring the SN Ia rate can put constraints on the evolutionary scenarios of different SN Ia progenitors. Methods.We used simulated data from the Dark Energy Science Collaboration (DESC) Data Challenge 2 (DC2) and LSST Data Preview 0 to measure the SN Ia rate on a 15 deg 2 region of the “wide-fast-deep” area. We selected a sample of SN candidates detected in difference images, associated them to the host galaxy with a specially developed algorithm, and retrieved their photometric redshifts. We then tested different light-curve classification methods, with and without redshift priors (albeit ignoring contamination from other transients, as DC2 contains only SN Ia). We discuss how the distribution in redshift measured for the SN candidates changes according to the selected host galaxy and redshift estimate. Results.We measured the SN Ia rate, analyzing the impact of uncertainties due to photometric redshift, host-galaxy association and classification on the distribution in redshift of the starting sample. We find that we are missing 17% of the SN Ia, on average, with respect to the simulated sample. As 10% of the mismatch is due to the uncertainty on the photometric redshift alone (which also affects classification when used as a prior), we conclude that this parameter is the major source of uncertainty. We discuss possible reduction of the errors in the measurement of the SN Ia rate, including synergies with other surveys, which may help us to use the rate to discriminate different progenitor models.

Astronomy & Astrophysics↗

A Standardized Analysis Process Using Digital Image Correlation to Calculate In Situ Cladding Strain from Modified Burst Tests for Fuel Performance Code Validation

Historical data collection on nuclear fuel cladding materials has focused on generating a statistically significant amount of data to assess the material and its failure behavior. Furthermore, data generated to support material model and failure criteria development were previously posttest evaluations, so a large number of tests was required to gain new understanding. A way to expedite this process is to develop techniques capable of generating large, high-fidelity data sets from a single test with lower uncertainty or quantified uncertainty. One such example of this approach is Oak Ridge National Laboratory’s use of modified burst tests (MBTs) to analyze the mechanical behavior and failure conditions of cladding during a simulated reactivity-initiated accident (RIA). Each test incorporates digital image correlation (DIC) analysis techniques that are used to assess the accumulated strain in situ, as well as eventual cladding failure. This work has been fruitful in defining strain-to-failure conditions for materials like silicon carbide (SiC) fiber–reinforced/SiC matrix composite tubes (SiC/SiC), iron-chromium-aluminum (FeCrAl) alloy tubes, and chromium-coated Zircaloy-4 tubes. However, there are numerous DIC software available, including open-source and proprietary software. The different DIC software use various algorithms to process images and calculate displacement values. Using these different software and algorithms can lead to varying results, and perhaps larger-than-expected uncertainties. In the present study, previously published MBT data encompassing a variety of test conditions were reanalyzed with two different DIC software to assess the variance in the calculated strain results. The data consisted of SiC/SiC, FeCrAl, and chromium-coated Zircaloy-4 tubes. Plots of the calculated strains during the transient revealed good agreement between the two DIC software. The average root-mean-square errors between the two software was 0.20% strain, which is slightly larger than a previously reported error value for these tests. In conclusion, this variance in results is low enough that this analysis method can be used for code validation.

Reactivity-initiated accident↗

Rapid monitoring of fermentations: a feasibility study on biological 2,3-butanediol production

2,3-butanediol (2,3-BDO) is an economically important platform chemical that can be produced by the fermentation of sugars using an engineered strain of Zymomonas mobilis . These fermentations require continuous monitoring and modification of fermentation conditions to maximize 2,3-BDO yields and minimize the production of the undesired coproducts glycerol and acetoin. Because of the time required for sampling and off-line chromatographic measurement of fermentation samples, the ability of fermentation scientists to modify fermentation conditions in a timely manner is limited. The goal of this study was to test if near-infrared spectroscopy (NIRS) along with multivariate statistics could reduce the time needed for this analysis and enable real-time monitoring and control of the fermentation. In this work we developed partial least squares (PLS) calibration models to predict the concentrations of glucose, xylose, 2,3-BDO, acetoin, and glycerol in fermentations via NIRS using two different spectrometers and two different spectroscopy modalities. We first evaluated the feasibility of rapid NIRS monitoring through experiments where we measured the signals from each analyte of interest and built NIRS-based PLS models using spectra from synthetic samples containing uncorrelated concentrations of these analytes. All analytes showed unique spectral signatures, and this initial modeling showed that all analytes could be detected simultaneously. We then began work with samples from laboratory fermentation experiments and tested the feasibility of regression model development across two spectral collection modalities (at-line and on-line) and two instruments: a laboratory-grade instrument and a low-cost instrument with a more limited spectral range. All modalities showed promise in the ability to monitor Z. mobilis fermentations of glucose and xylose to 2,3-BDO. The low-cost instrument displayed a lower signal-to-noise ratio than the laboratory-grade instrument, which led to comparatively lower performance overall, but still provided sufficient accuracy to monitor fermentation trends. While the ease of use of on-line monitoring systems was favored as compared to at-line systems due to the lack of sampling required and potential for automated process control, we observed some decrease in performance due to the additional complexity of the sample matrix. We have demonstrated that NIRS combined with multivariate analysis can be used for at-line and on-line monitoring of the concentrations of glucose, xylose, 2,3-BDO, acetoin, and glycerol during Z. mobilis fermentations. The decrease in signal-to-noise ratio when using a low-cost spectrometer led to greater prediction error than the laboratory-grade spectrometer for at-line monitoring. The on-line monitoring modality showed great promise for real time process control via NIRS.

09 BIOMASS FUELS↗

Comparison of CERES SYN1deg Radiative Fluxes with Those Derived from Observations at the ARM ENA Site

Profiles of radiative fluxes simulated from thermodynamic and cloud observations made at the Atmospheric Radiation Measurement (ARM) eastern North Atlantic (ENA) site for a 6-yr period are termed as ENARad. ENARad radiative fluxes are compared to those from the Clouds and the Earth’s Radiant Energy System (CERES) 1°-resolution synoptic product (SYN1deg)-simulated radiative flux profiles as well as the CERES instrument observed top-of-the-atmosphere (TOA) fluxes and ground site broadband radiometer measurements. Monthly average differences between ENARad and surface radiometer reported fluxes and differences between ENARad, SYN1deg, and observed fluxes at TOA were statistically insignificant. SYN1deg significantly overestimated surface downwelling shortwave flux by 12 ± 52 W m −2 and surface downwelling longwave flux by 5 ± 20 W m −2 on monthly time scales. Such overestimations were traced to a moister and warmer subcloud layer, a drier cloud layer, and a moister and colder above-cloud-free troposphere in the ancillary thermodynamic and cloud properties used by SYN1deg than observed. Similarly, low-cloud coverage, boundaries, and liquid water paths utilized by SYN1deg were also significantly higher than observed. Intramodel differences in the hourly values of shortwave fluxes exceeded 100 W m−2 at the TOA and the surface. These differences were also due to inaccuracies in the representation of low-cloud properties within the SYN1deg product relative to those determined by ENA ARM instrumentation and used as ENARad ancillary data. Results presented are relevant to investigations employing the CERES SYN1deg data product, studies that estimate radiative fluxes from surface-based or satellite-borne observations, and comparative analyses of radiative fluxes derived using different methodological approaches.

54 ENVIRONMENTAL SCIENCES↗

Generating mock galaxy catalogues for flux-limited samples like the DESI Bright Galaxy Survey

ABSTRACT Accurate mock galaxy catalogues are crucial to validate analysis pipelines used to constrain dark energy models. We present a fast HOD-fitting method which we apply to the AbacusSummit simulations to create a set of mock catalogues for the DESI Bright Galaxy Survey, which contain r-band magnitudes and $(g-r)$ colours. The halo tabulation method fits HODs for different absolute magnitude threshold samples simultaneously, preventing unphysical HOD crossing between samples. We validate the HOD fitting procedure by fitting to real-space clustering measurements and galaxy number densities from the MXXL BGS mock, which was tuned to the SDSS and GAMA surveys. The best-fitting clustering measurements and number densities are mostly within the assumed errors, but the clustering for the faint samples is low on large scales. The best-fitting HOD parameters are robust when fitting to simulations with different realizations of the initial conditions. When varying the cosmology, trends are seen as a function of each cosmological parameter. We use the best-fitting HOD parameters to create cubic box and cut sky mocks from the AbacusSummit simulations, in a range of cosmologies. As an illustration, we compare the ${}^{0.1}M_r\lt -20$ sample of galaxies in the mock with BGS measurements from the DESI one-percent survey. We find good agreement in the number densities, and the projected correlation function is reasonable, with differences that can be improved in the future by fitting directly to BGS clustering measurements. The cubic box and cut-sky mocks in different cosmologies are made publicly available.

79 ASTRONOMY AND ASTROPHYSICS↗

Future Intensity‐Duration‐Frequency Curves of Extreme Precipitation in the Midwest United States From Convection‐Permitting Modeling

Abstract During the last four decades, global warming has statistically significant intensified extreme precipitation events in the Midwestern United States (defined here as the region covering Illinois, Indiana, Ohio, and Kentucky), leading to increased risks to human life, property, and infrastructure. To enable climate change adaptation and resilience across various economic and social sectors in this region, updated information about future climate changes, specifically at finer spatial scales, is essential. Leveraging a new 150‐year dynamical downscaling data set at convection‐permitting resolution, this study introduces a framework to construct the projected future intensity‐duration‐frequency (IDF) curves of heavy precipitation, which are prominent tools for infrastructure design and water resources management. This framework generates IDF curves at both sub‐daily and multi‐day duration utilizing hourly in situ observations as well as quantile‐based statistical techniques in bias‐correction and return levels selection. The assumption of non‐stationarity in the distribution parameter fitting process is also implemented in this workflow. Compared to historical IDF curves for 1980–2022, future projected IDF curves for 2058–2100 under Representative Concentration Pathway (RCP) 4.5 and RCP 8.5 scenarios indicate an average intensity increase of approximately 15% and 25%, respectively, across 74 stations, considering both annual and seasonal timescales. Future projections suggest that extreme precipitation events may become more severe across six investigated return periods, with longer return periods showing a greater increase. The frequency of future extreme precipitation events in the Midwest region is also projected to double. Furthermore, current results reveal spatial heterogeneity of future trends across stations owing to the high‐resolution input data set. Plain Language Summary This study investigates the evolving nature of extreme precipitation events in the Midwestern United States under a changing climate. By leveraging a high‐resolution dynamical downscaling data set, we construct projected intensity‐duration‐frequency (IDF) curves for future extreme rainfall events. These curves serve as vital tools for infrastructure planning and water resource management. Our analysis reveals a significant increase in both the intensity and frequency of extreme precipitation events in the region. Future projected IDF curves for the late century indicate an average intensity increase of approximately 15%–25% compared to historical values. Moreover, the frequency of such events is expected to double. Spatial heterogeneity in future trends is observed across different stations within the Midwest, highlighting the importance of high‐resolution modeling in capturing localized climate variability. These findings underscore the urgent need for climate adaptation strategies to mitigate the increasing risks associated with extreme precipitation events in the region. Key Points This study introduces a workflow to construct future intensity‐duration‐frequency (IDF) curves over the Midwest United States using a new convection‐permitting modeling data set The current IDF construction workflow reproduces well the historical observed IDF 30 curves in summer months with median relative errors of 2.4% among 74 stations and 6 investigated durations The projected IDF curves show diverse future trends of extreme precipitation across stations, with intensity increases of approximately 15% and 25% under RCP4.5 and RCP8.5 climate scenarios, respectively, and a doubling of frequency on average

Nguyen, Trung↗

From chiral effective field theory to perturbative QCD: A Bayesian model mixing approach to symmetric nuclear matter

Constraining the equation of state (EOS) of strongly interacting, dense matter is the focus of intense experimental, observational, and theoretical effort. Chiral effective field theory (𝜒⁢EFT ) can describe the EOS between the typical densities of nuclei and those in the outer cores of neutron stars, while perturbative QCD (pQCD) can be applied to properties of deconfined quark matter, both with quantified theoretical uncertainties. However, describing the full range of densities in between with a single EOS that has well-quantified uncertainties is a challenging problem. Bayesian multimodel inference from 𝜒⁢EFT and pQCD can help bridge the gap between the two theories. In this work, we introduce a correlated Bayesian model mixing framework that uses a Gaussian process (GP) to assimilate different information into a single QCD EOS for symmetric nuclear matter. The present implementation uses a stationary GP to infer this mixed EOS solely from the EOSs of 𝜒⁢EFT and pQCD while accounting for the truncation errors of each theory. The GP is trained on the pressure as a function of number density in the low- and high-density regions where 𝜒⁢EFT and pQCD are, respectively, valid. We impose priors on the GP kernel hyperparameters to suppress unphysical correlations between these regimes. This, together with the assumption of stationarity, results in smooth 𝜒⁢EFT-to-pQCD curves for both the pressure and the speed of sound. We show that using uncorrelated mixing requires uncontrolled extrapolation of at least one of 𝜒⁢EFT or pQCD into regions where the perturbative series breaks down and leads to an acausal EOS. Here, we also discuss extensions of this framework to nonstationary and less differentiable GP kernels, its future application to neutron-star matter, and the incorporation of additional constraints from nuclear theory, experiment, and multimessenger astronomy.

Bayesian methods↗

Batch VUV4 characterization for the SBC-LAr10 scintillating bubble chamber

The Scintillating Bubble Chamber (SBC) collaboration purchased 32 Hamamatsu VUV4 silicon photomultipliers (SiPMs) for use in SBC-LAr10, a bubble chamber containing 10 kg of liquid argon. A dark-count characterization technique, which avoids the use of a single-photon source, was used at two temperatures to measure the VUV4 SiPMs breakdown voltage (V BD ), the SiPM gain (g SiPM ), the rate of change of g SiPM with respect to voltage (m), the dark count rate (DCR), and the probability of a correlated avalanche (P CA ) as well as the temperature coefficients of these parameters. A Peltier-based chilled vacuum chamber was developed at Queen's University to cool down the Quads to 233.15 ± 0.2 K and 255.15 ± 0.2 K with average stability of ±20 mK. An analysis framework was developed to estimate V BD to tens of mV precision and DCR close to Poissonian error. The temperature dependence of V BD was found to be 56 ± 2 mV K -1 , and m on average across all Quads was found to be (459 ± 3(stat.)±23(sys.))× 10 3 e- PE -1 V -1 . The average DCR temperature coefficient was estimated to be 0.099 ± 0.008 K -1 corresponding to a reduction factor of 7 for every 20 K drop in temperature. The average temperature dependence of P CA was estimated to be 4000 ± 1000 ppm K -1 . P CA estimated from the average across all SiPMs is a better estimator than the P CA calculated from individual SiPMs, for all of the other parameters, the opposite is true. All the estimated parameters were measured to the precision required for SBC-LAr10, and the Quads will be used in conditions to optimize the signal-to-noise ratio.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Videos, photos, and AI-derived grain size data associated with “High-throughput AI Video Surveys Enable Reproducible Multiscale Sediment Size Mapping, with Implications for Hydrobiogeochemical Parameterization”

NOTE: The manuscript associated with this data package is currently in review. The data may be revised based on reviewer feedback. Upon manuscript acceptance, this data package will be updated with the final dataset and additional metadata. This data package is associated with the manuscript “High-throughput AI Video Surveys Enable Reproducible Multiscale Sediment Size Mapping, with Implications for Hydrobiogeochemical Parameterization” under review. This data package includes five data types: 1) raw photos and videos from drone survey and walking smartphone surveys; 2) images derived from raw videos; 3) manual labeling of reference scales; 4) metadata for all images and photo resolution derived from artificial intelligence (AI) models or manual labels, 5) grain size data obtained from AI models for all photos, 6) metadata and grain size data after quality control, 7) summaries of sample efficiency for all data, and 8) computational fluid dynamics (CFD) data used to support hydro-biogeochemical (HBGC) parameter estimation. Such data is used to 1) demonstrate significant improvements in accuracy, efficiency, and quality control for grain size data collection with the help of AI models, 2) study the spatial heterogeneity of grain size and observation reproducibility based on tens of thousands of data points generated by the AI models, and 3) evaluate the impacts of grain size heterogeneity on key HBGC parameters across sediment-to-reach and hourly-to-yearly scales. In particular, the data package contains 116 folders and 179696 files. The files include 41 videos in .mov format, 64047 photos in .jpg format, 13541 video-derived photos in .png format, 12747 segmentation mask data in .tif format, 12747 segmentation data in .json format, 24771 .csv files that with metadata and grain size for each individual photo as well as water depth and velocity data from CFD and observation, 51791 .txt files of raw AI predicted labels, and 11 flight record data in .srt format. The summary for all metadata and grain size statistics information is included in “Scales_V3_NG.csv” and “Statistics_V3_NG.csv”. The summary for data that pass data quality control (QC) level 0-2 is included in “QCStatistics_V3_NG.csv”. The QC level 0 represents photos whose photo resolution is positive, excluding photos that miss reference scale. The QC level 1 means reference scale circularity uncertainty is less than 5% for smartphone images while representing photo resolution is larger than 0.44 mm/pixel for drone images. The QC level 2 means excluding photos whose grain number is less than 100, a minimum number of grains recommended by classic literature. The summary for each video’s name, length, frame rates, survey area, grain number, survey efficiency, etc. can be found in “QCSummary_V3_NG.csv”. The summary for site name, GPS coordinates, and number of images at each site can be found in “SitesSummary_V3_*.csv” files. Overall computational efficiency summary is reported in Table 4 of accompanying manuscript. Additionally, the nitrate concentration data used in this work was downloaded from an existing dataset published on ESS-DIVE (Boat-Dragged Sensor Hanford Reach.csv; Conner A. et al., 2020). We thank the United States Forest Service, Washington Department of Fish and Wildlife, Washington Department of Natural Resources, Cowiche Canyon Conservatory, Port of Benton, and the Confederated Tribes and Bands of the Yakama Nation for access to field locations where the data were collected. We also thank the Yakama Nation Tribal Council and Yakama Nation Fisheries for working with us to facilitate data collection and optimization of data usage according to their values and worldview.

54 ENVIRONMENTAL SCIENCES↗

Statistical Correlation of Heliostat Pointing Deviation With Wind

This work was carried out as part of the Heliostat Consortium (HelioCon) Field Deployment subtask with the aim to develop a reduced order model framework for correlating wind speed and pointing deviation of a heliostat facet. There are only sparse field measurements of heliostat pointing deviations and accompanying wind conditions published in the literature. Heliostat test standards, such as IEC 62862-4-3, propose a suite of tests including laser pointing repeatability at wind speeds below 4 m/s, and provide technical requirements for heliostat slope and tracking deviations in coarse average wind speed bins of 4 m/s, 6 m/s, and 8 m/s. In addressing the gap of the variation of heliostat pointing deviation with wind speed, field measurements of laser pointing on a grid target and wind conditions were analyzed in this study at the Third-Party Metrology Platform at the National Laboratory of the Rockies (NLR) Flatirons Campus. Horizontal pointing deviations were found to follow a logarithmic relationship with peak wind speed, whereas vertical pointing deviations follow an exponential relationship with peak wind speed. Both horizontal and vertical pointing deviations also follow a second order polynomial relationship, as expected from the proportionality of elastic loads and deformations with the square of wind speed. The results indicate that heliostat facet pointing deviations in the vertical direction increase at a faster rate than in the horizontal direction with increasing wind speed over the tested range, however these are dependent on the heliostat structural design. Next steps are recommended for additional field measurements to confirm a linear relationship of pointing deviation with applied moment on a heliostat facet, and to distinguish between gravity-induced and wind-induced pointing deviations at different elevation angles. The derived correlations in the preliminary analysis in this report serve as a case study for heliostat developers and plant operators to estimate the wind-induced pointing deviations and their variation with peak gust wind speed. Next steps in future work would recommend higher resolution and longer duration datasets for different elevation angles and wind directions to reduce uncertainties and variance of collected laser beam spot data and their correlations with bin-averaged wind speed.

17 WIND ENERGY↗