Search NASASearch

SEARCH · Search NASA

Results for “Data Interpretation, Statistical”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Using the optimal combined index weight ratio to improve the probability of anomaly detection in big area additive manufacturing

Big Area Additive Manufacturing (BAAM) of composites requires significant time, energy, and material, so it is critical to reduce production inefficiencies to make functional parts without multiple iterations. Statistical process control coupled with Principal Component Analysis (PCA) is a powerful technique that provides a quick, computationally inexpensive, and intuitive way for operators to detect defects that form in a manufacturing process without massive datasets. Recently, a combined index that is a weighted sum of the Hotelling's T 2 and squared residual error statistics has been proposed that can be monitored in one chart, improving interpretation accuracy and simplicity. However, the literature does not offer a formal method to optimise the weights. Here, we introduce two new approaches to the traditional weight selection approach using simulated and BAAM image data. Approach 1 uses a theoretically motivated optimum inspired by probabilistic principal component analysis. Approach 2 systematically varies the ratio of the weights to find the optimum. We show that approach 1 delivers optimal anomaly detection performance in select cases while approach 2 fares better in practice. Surprisingly, we also show that choosing a more complex PCA model has a minimal negative impact on anomaly detection performance compared to a more simplistic model.

3-dimensional printing

Q -score as a reliability measure for protein, nucleic acid and small-molecule atomic coordinate models derived from 3DEM maps

Atomic coordinate models are important for the interpretation of 3D maps produced with cryoEM and cryoET (3D electron microscopy; 3DEM). In addition to visual inspection of such maps and models, quantitative metrics can inform about the reliability of the atomic coordinates, in particular how well the model is supported by the experimentally determined 3DEM map. A recently introduced metric, Q-score, was shown to correlate well with the reported resolution of the map for well fitted models. Here, we present new statistical analyses of Q-score based on its application to ∼10 000 maps and models archived in the EMDB (Electron Microscopy Data Bank) and PDB (Protein Data Bank). Further, we introduce two new metrics based on Q-score to represent each map and model relative to all entries in the EMDB and those with similar resolution. We explore through illustrative examples of proteins, nucleic acids and small molecules how Q-scores can indicate whether the atomic coordinates are well fitted to 3DEM maps and also whether some parts of a map may be poorly resolved due to factors such as molecular flexibility, radiation damage and/or conformational heterogeneity. These examples and statistical analyses provide a basis for how Q-scores can be interpreted effectively in order to evaluate 3DEM maps and atomic coordinate models prior to publication and archiving.

B factors

Data-driven high-dimensional statistical inference with generative models

Crucial to many measurements at the LHC is the use of correlated multi-dimensional information to distinguish rare processes from large backgrounds, which is complicated by the poor modeling of many of the crucial backgrounds in Monte Carlo simulations. In this work, we introduce HI-SIGMA, a method to perform unbinned high-dimensional statistical inference with data-driven background distributions. In contradistinction to many applications of Simulation Based Inference in High Energy Physics, HI-SIGMA relies on generative ML models, rather than classifiers, to learn the signal and background distributions in the high-dimensional space. These ML models allow for interpretable inference while also incorporating model errors and other sources of systematic uncertainties. We showcase this methodology on a simplified version of a di-Higgs measurement in the bbγγ final state, where the di-photon resonance allows for background interpolation from sidebands into the signal region. We demonstrate that HI-SIGMA provides improved sensitivity as compared to standard classifier-based methods, and that systematic uncertainties can be straightforwardly incorporated by extending methods which have been used for histogram based analyses.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS

pyTCR: A tropical cyclone rainfall model for python

pyTCR is a climatology software package developed in the Python programming language. It integrates the capabilities of several legacy physical models and increases computational efficiency to allow rapid estimation of tropical cyclone (TC) rainfall consistent with the large-scale environment. Specifically, pyTCR implements a horizontally distributed and vertically integrated model [Zhu et al., 2013] for simulating rainfall driven by TCs. Along storm tracks, rainfall is estimated by computing the cross-boundary-layer, upward water vapor transport caused by different mechanisms including frictional convergence, vortex stretching, large-scale baroclinic effect (i.e., wind shear), topographic forcing, and radiative cooling [Lu et al., 2018]. The package provides essential functionalities for modeling and interpreting spatio-temporal TC rainfall data. pyTCR requires a limited number of model input parameters, making it a convenient and useful tool for analyzing rainfall mechanisms driven by TCs. To sample rare (most intense) rainfall events that are often of great societal interest, pyTCR adapts and leverages outputs from a statistical-dynamical TC downscaling model [Lin et al., 2023] capable of rapidly generating a large number of synthetic TCs given a certain climate. As a result, pyTCR significantly reduces computational effort and improves the efficiency in capturing extreme TC rainfall events at the tail of the distributions from limited datasets. Furthermore, the TC downscaling model is forced entirely by large-scale environmental conditions from reanalysis data or coupled General Circulation Models (GCMs), simplifying the projection of TC-induced rainfall and wind speed under future climate using pyTCR. Finally, pyTCR can be coupled with hydrological and wind models to assess risks associated with independent and compound events (e.g., storm surges and freshwater flooding).

54 ENVIRONMENTAL SCIENCES

Search for dark matter produced in association with a Higgs boson decaying to a τ lepton pair in proton-proton collisions at $\sqrt{s}=13$ TeV

A search for dark matter particles produced in association with a Higgs boson decaying into a pair of τ leptons is performed using data collected in proton-proton collisions at a center-of-mass energy of 13 TeV with the CMS detector. The analysis is based on a data set corresponding to an integrated luminosity of 101 fb −1 collected in 2017–2018. No significant excess over the expected standard model background is observed. This result is interpreted within the frameworks of the 2HDM+a and baryonic Z′ benchmark simplified models. The 2HDM+a model is a type-II two-Higgs-doublet model featuring a heavy pseudoscalar with an additional light pseudoscalar. Upper limits at 95% confidence level are set on the product of the production cross section and the branching fraction for each of these two simplified models. Heavy pseudoscalar boson masses between 400 and 700 GeV are excluded for a light pseudoscalar mass of 100 GeV. For the baryonic Z′ model, a statistical combination is made with an earlier search based on a data set of 36 fb −1 collected in 2016. In this model, Z′ boson masses up to 1050 GeV are excluded for a dark matter particle mass of 1 GeV.

Dark Matter

Signal-preserving CMB component separation with machine learning

Analysis of microwave sky signals, such as the cosmic microwave background, often requires component separation using multifrequency methods, whereby different signals are isolated according to their different frequency behaviors. Many so-called blind methods, such as the internal linear combination (ILC), make minimal assumptions about the spatial distribution of the signal or contaminants, and only assume knowledge of the frequency dependence of the signal. The ILC produces a minimum-variance linear combination of the measured frequency maps. In the case of Gaussian, statistically isotropic fields, this is the optimal linear combination, as the variance is the only statistic of interest. However, in many cases the signal we wish to isolate, or the foregrounds we wish to remove, are non-Gaussian and/or statistically anisotropic (in particular for the case of Galactic foregrounds). In such cases, it is possible that machine learning (ML) techniques can be used to exploit the non-Gaussian features of the foregrounds and thereby improve component separation. However, many ML techniques require the use of complex, difficult-to-interpret operations on the data. We propose a hybrid method whereby we train an ML model using only combinations of the data that , and combine the resulting ML-predicted foreground estimate with the ILC solution to reduce the error from the ILC. We demonstrate our methods on simulations of extragalactic temperature and Galactic polarization foregrounds and show that our ML model can exploit non-Gaussian features, such as point sources and spatially varying spectral indices, to produce lower-variance maps than ILC—e.g., reducing the variance of the B-mode residual by factors of up to 5—while preserving the signal of interest in an unbiased manner. Moreover, we often find improved performance even when applying our ML technique to foreground models on which it was not trained. Published by the American Physical Society 2025

McCarthy, Fiona (ORCID:0000000253893565)

Dark Energy Survey: DESI-independent angular BAO measurement

In this work, we present a measurement of the angular baryon acoustic oscillation (BAO) scale from the completed Dark Energy Survey (DES) dataset excluding the area of overlap with the Dark Energy Spectroscopic Instrument (DESI). We follow the same methodology and validation process as in the DES Y6 BAO analysis. We interpret the impact of this measurement in the context of the statistical preference for 𝑤 0 ⁢𝑤 𝑎 cold dark matter (CDM) over Λ⁢CDM when combined with DES Y5 Type Ia supernovae (SN), Planck CMB, and DESI BAO. Based on our previous work, using the full Y6 DES BAO sample, in combination with SN, CMB and DESI data release 1 (DR1) BAO, added 0.3⁢𝜎 in this preference (from 3.7⁢𝜎 to 4.0⁢𝜎), but this ignored possible correlations between datasets. Using our new DESI-independent DES BAO likelihood instead, we find a smaller increase in the statistical preference for 𝑤 0 ⁢𝑤 𝑎 ⁢CDM, from 3.7⁢𝜎 to 3.8⁢𝜎 when using DESI DR1 BAO, and from 4.0⁢𝜎 to 4.1⁢𝜎 when updating to the more recent DESI data release 2 (DR2) BAO. These significances reduce to 3.1⁢𝜎 when using the new calibrated DES SN-Dovekie. Alongside this work, we publicly release baofit_wtheta, the BAO fitting code for the angular correlation function used in the DES Y6 BAO analysis.

79 ASTRONOMY AND ASTROPHYSICS

Level structure of light neutron-rich La isotopes beyond the 𝑁 = 82 shell closure

Here, the high spin excited states of Lanthanum isotopes 140–143 La, above the 𝑁 = 82 closed shell, were populated in fission reactions. The prompt 𝛾-ray transitions were measured using two complementary methods: (a) in coincidence with the isotopically identified fragments produced in the fission of the 238 U + 9 Be system using the Variable Mode Spectrometer (VAMOS++) and the Advanced Gamma Tracking Array (AGATA) spectrometer, and (b) high statistics threefold 𝛾−𝛾−𝛾 and fourfold 𝛾−𝛾−𝛾−𝛾 coincidence data from the spontaneous fission of 252 Cf using the Gammasphere. This work reports the first identification of a pair of parity doublet structures in 143 La and the new high spin level structure in 140–142 La from prompt 𝛾-ray spectroscopy. The level structures are interpreted in terms of the systematics of neighboring odd-𝑍 nuclei above the 𝑍 = 50 shell closure and large-scale shell model calculations. The present results indicate the presence of stable octupole deformation in 143 La. The excitation energy pattern and their comparison with neighboring isotones, moving away from the 𝑁 = 82 closed shell, point towards a transition from single-particle structures to an alternating parity rotational band structure in the La isotopic chain.

Navin, A. [Centre National de la Recherche Scienti

CAHS: Context-Aware Homology Search

Protein homology search is foundational to bioinformatics: it supports annotation transfer, structure/function inference, and evolutionary analysis over rapidly expanding sequence repositories (e.g., UniProtKB). Profile hidden Markov models (pHMMs), as implemented in HMMER, remain the most widely trusted approach because they provide statistically calibrated E-values; however, their gap behavior is fixed once a profile is trained, despite biological evidence that insertion/deletion tolerance varies across flexible loops and intrinsically disordered regions. We present CAHS (Context-Aware Homology Search), a lightweight query-time adapter for pHMM search that incorporates learned and biologically motivated signals without changing HMMER's downstream search pipeline or its calibrated E-value reporting. Given a query sequence, CAHS computes per-residue representations from a protein language model and a disorder predictor, maps these to profile coordinates, and modulates only match-state transition rows (gap-open and gap-extension probabilities) while preserving Plan7 constraints. We comprehensively evaluate CAHS across six structurally diverse protein families and multi-domain architectures against a 570k-sequence target corpus. CAHS expands detection capability, retrieving thousands of additional remote homologs at relaxed thresholds by maintaining alignment quality through flexible regions. For multi-domain proteins, context-aware modulation resolves 94% of fragmented alignments. Crucially, CAHS preserves hit-set invariance at stringent operating points (E<10-10), demonstrating increased statistical confidence without inflating false positives. Furthermore, sharper statistical distinction between homologs and background noise during early filter stages yields up to a 3.87× acceleration in end-to-end wall-clock time on high-performance computing clusters. Overall, CAHS illustrates a practical AI-for-science design pattern: augmenting a trusted probabilistic model with query-specific learned signals to improve interpretable, reproducible inference in data-rich biology.

Bhattaram, Swethasree [Georgia Institute of Techno

Comparing multi-source urban flood indicators: satellite, simulation, and citizen-reported data

Urban flooding arises from complex mechanisms, making it challenging to capture accurately with a single detection method. This study evaluates three complementary approaches to detect flooding across three Chicago neighborhoods: (i) Sentinel-1 synthetic aperture radar (SAR), offering weather-independent, high-resolution (10 m) imagery of surface inundation; (ii) the storm water management model (SWMM), simulating combined sewer overflow and drainage performance; and (iii) citizen-generated 311 service requests, capturing observed flooding impacts. By analyzing six storms ranging from severe to mild, we examine how each source uniquely contributes to identifying urban flood events. SAR imagery effectively identifies standing water but can miss brief flooding due to satellite revisit constraints. SWMM provides detailed insights into system-wide drainage behavior yet may underestimate localized street-level flooding. Meanwhile, 311 calls reflect real-world flooding impacts but are vulnerable to underreporting. Statistical overlap analysis highlights chronic flood hotspots repeatedly identified across multiple detection methods, indicating persistent infrastructure and topographic vulnerabilities. Temporal analysis further reveals that while SWMM flooding aligns closely with rainfall peaks, 311 calls typically precede or persist beyond these peaks. Our findings emphasize the value of using satellite observations, hydrological modeling, and resident-reported data in a complementary manner to better interpret patterns in flood timing, severity, and spatial distribution—providing insights that can inform targeted infrastructure improvements and contribute to urban flood resilience planning.

311

High-count-rate effects in event processing for the XRISM/Resolve X-ray microcalorimeter. II. Energy scale and resolution in orbit

The Resolve instrument on the X-ray Imaging and Spectroscopy Mission (XRISM) uses a 36 pixel microcalorimeter designed to deliver high-resolution, non-dispersive X-ray spectroscopy. Although it is optimized for extended sources with low count rates, Resolve observations of bright point sources are still able to provide unique insights into the physics of these objects, as long as high-count-rate effects are addressed in the analysis. These effects include the loss of exposure time for each pixel, changes in the energy scale, and changes in the energy resolution. To investigate these effects under realistic observational conditions, we observed the bright X-ray source, the Crab Nebula, with XRISM at several offset positions with respect to the Resolve field of view and with continuous illumination from 55 Fe sources on the filter wheel. For the spectral analysis, we excluded data where exposure-time loss was too significant to ensure reliable spectral statistics. The energy scale at 6 keV shows a slight negative shift in the high-count-rate regime. The energy resolution at 6 keV worsens as the count rate in electrically neighboring pixels increases, but can be restored by applying a nearest-neighbor coincidence cut (“cross-talk cut”). We examined how these effects influence the observation of bright point sources, using GX 13+1 as a test case, and identified an eV-scale energy offset at 6 keV between the inner (brighter) and outer (fainter) pixels. Users who seek to analyze velocity structures on the order of tens of km s–1 should account for such high-count-rate effects. These findings will aid in the interpretation of Resolve data from bright sources and provide valuable considerations for designing and planning for future microcalorimeter missions.

X-rays: general

Temporal Convolutional Network Using Empirical Mode Decomposition to Detect Faults in Grid Connected Systems

Grid-connected power electronic systems require timely and reliable fault detection to prevent equipment damage and reduce downtime. This paper presents a forecasting-based anomaly detection pipeline that decomposes voltage and current measurements into intrinsic mode functions (IMFs) using empirical mode decomposition (EMD), then trains a causal temporal convolutional network (TCN) on normal-operation IMF data to predict short-horizon future dynamics. Deviations between forecasts and observations are summarized as reliability-weighted residual scores and thresholded per sensor using robust statistics with temporal persistence constraints to suppress false positives. To reduce runtime, EMD is performed on downsampled signals for detection, while raw-rate EMD is applied only within a short region of interest for high-frequency interpretability near detected events. Results on a simulated grid-connected converter system demonstrate that IMF-domain forecasting improves anomaly separability relative to raw-signal forecasting and provides interpretable evidence of faults across decomposition channels.

Sutton, Elizabeth [ORNL] (ORCID:0009000078885935)

Meteoric 10Be Flux Calibration Data for the East River Watershed, Colorado, USA

This data package contains tabular and geospatial data used to quantify and model meteoric beryllium-10 fluxes in the East River watershed, Colorado, USA. The tabular component includes calibration-site data from five glacial moraine sites and includes environmental variables used to evaluate spatial controls on meteoric 10Be delivery, including elevation, mean annual precipitation (MAP), mean snow depth, and mean snow water equivalent (SWE). These site-level data were used to compare observed fluxes with environmental gradients across the watershed and to evaluate the effects of erosion correction on flux estimates. The package also includes supporting slope and curvature values used to assess topographic inputs to the erosion analysis. A second component of the data package contains updated manuscript tables and regression outputs used to summarize the relationships between meteoric 10Be flux and environmental predictors. These tables include meteoric 10Be sample information and AMS results, site-level environmental values, site-level meteoric 10Be inventory and flux values, watershed-averaged predicted fluxes, soil bulk density measurements, fine-fraction values, soil pH measurements, and regression statistics including slope, intercept, coefficient of determination, and p-value. The regression products include both standard linear regressions and regressions in which the intercept is constrained to pass through zero, and they support the analyses presented in the companion manuscript. Together, these tabular files provide the numerical basis for the manuscript tables and the regression-based interpretation of meteoric 10Be flux variability in a snow-dominated mountain watershed. The geospatial component of the package consists of GeoTIFF raster files used to generate the map products presented in Figures 2 and 6 of the companion manuscript. These rasters represent watershed-scale spatial layers for environmental variables and regression-based predictions of meteoric 10Be flux. This dataset contains comma-separated values files (.csv), Microsoft Excel files (.xlsx), GeoTIFF raster files (.tif), and upporting metadata files, including CSV data dictionaries and readme text files (.csv, .txt). The tabular files can be opened with standard spreadsheet software, and the raster files can be viewed and analyzed in GIS software such as ArcGIS Pro or QGIS. Together, these files document the numerical and spatial datasets used to calibrate and predict meteoric 10Be delivery in the East River watershed.

East River

Visual Analytics of Crosstalk in Quantum Hardware

Crosstalk remains a major obstacle to building scalable and fault-tolerant quantum computers. Conventional diagnostic techniques-often based on numerical simulation or statistical modeling-struggle to scale with hardware complexity and offer limited interpretability. In this work, we present a visual analytics framework for diagnosing qubit crosstalk using lightweight, circuit-based models integrated with an interactive user interface. Our approach quantifies correlations between active and idle qubits under parameterized single- and twoqubit operations, enabling detection of both spatial and gateinduced crosstalk. The system incorporates qubit topology and gate performance data to support sector-based exploration and correlation mapping. This tool assists users in identifying correlated error sources, informing qubit placement strategies, and guiding noise-aware circuit design.

Chae, Junghoon [ORNL] (ORCID:0000000206016746)

Probabilistic inference of the structure and orbit of Milky Way satellites with semi-analytic modelling

Semi-analytic modelling furnishes an efficient avenue for characterizing dark matter haloes associated with satellites of Milky Way-like systems, as it easily accounts for uncertainties arising from halo-to-halo variance, the orbital disruption of satellites, baryonic feedback, and the stellar-to-halo mass (SMHM) relation. We use the SatGen semi-analytic satellite generator, which incorporates both empirical models of the galaxy–halo connection as well as analytic prescriptions for the orbital evolution of these satellites after accretion onto a host to create large samples of Milky Way-like systems and their satellites. By selecting satellites in the sample that match observed properties of a particular dwarf galaxy, we can infer arbitrary properties of the satellite galaxy within the cold dark matter paradigm. For the Milky Way’s classical dwarfs, we provide inferred values (with associated uncertainties) for the maximum circular velocity v max and the radius r max at which it occurs, varying over two choices of baryonic feedback model and two prescriptions for the SMHM relation. While simple empirical scaling relations can recover the median inferred value for v max and r max , this approach provides realistic correlated uncertainties and aids interpretability. We also demonstrate how the internal properties of a satellite’s dark matter profile correlate with its orbit, and we show that it is difficult to reproduce observations of the Fornax dwarf without strong baryonic feedback. Furthermore, the technique developed in this work is flexible in its application of observational data and can leverage arbitrary information about the satellite galaxies to make inferences about their dark matter haloes and population statistics.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC

Geology of the One Earth Energy Site

The One Earth Energy site is one of two sites in the Illinois Storage Corridor (ISC) project. The objectives of the ISC project is to accelerate commercial deployment of carbon capture utilization and storage at two individual sites and receive approvals for Underground Injection Control (UIC) Class VI permits for construction at each site. At the One Earth Energy site, an extensive data collection program was undertaken, which included the drilling of a test well (One Earth Energy #1 [OEE #1]), four 2D seismic lines, and a small 3D seismic survey. The OEE #1 well was drilled in 2022 and acquired extensive core, log, and testing data to characterize the subsurface geology of the site. Coring was focused on the storage interval, the Mt. Simon Sandstone, and the confining interval, the Eau Claire Formation. The core and log data were used to evaluate the sedimentology and sequence stratigraphy, as well as to develop the conceptual geologic model. This report includes the geological summaries of the Mt. Simon Sandstone and the Eau Claire Formation. The extensive analysis of the log data is included in the petrophysical section, showing ranges of porosity, estimated pore size, and the mineral content of selected zones in the well. The separate petrographic technical report entitled “Petrographic and Advanced Geologic Characterization Report on One Earth Energy #1 (API# 1211325373)”, report number DOE-UIUC-0031892-04, details thin section point-counting analysis that includes mineralogical and pore space analysis, including grain size analysis, annotated thin section photomicrographs, scanning electron microscopy (SEM) with energy dispersive X-ray spectroscopy (EDS), and statistics of grain size analysis on Mt. Simon thin sections from OEE #1. The final OEE #1 well data to be included in this geology report is the routine core analysis of both whole core plugs and rotary sidewall core plugs. In addition to the OEE #1 well, four 2D seismic lines and a small 3D survey were acquired as part of the overall subsurface geological characterization. This geology report references the seismic interpretation report, entitled “One Earth Energy Site Seismic Interpretation Task 5.0”, report number DOE-UIUC-0031892-07. This report details the stratigraphic and structural interpretation of the 2D and 3D seismic data acquired at the One Earth Energy site. The 2D seismic data was acquired in 2019 and 2021, and the 3D survey was acquired in 2022. The objectives of the seismic programs were to contribute to the subsurface characterization of the Mt. Simon-Eau Claire Storage Complex by evaluating the continuity of potential storage reservoirs and containment intervals across the project area, and to determine if any geologic features are present that would increase containment risk to the proposed carbon storage project.

09 BIOMASS FUELS

Determination of proton PDF uncertainties with Markov chain Monte Carlo

We present an analysis of parton distribution functions (PDFs) of the proton using Markov chain Monte Carlo (MCMC) methods. The MCMC approach naturally implements Bayes’ theorem and, thus, provides a means to directly sample the underlying probability distribution—in this case, the probability distribution of the PDF parameters. This allows for a straightforward propagation of the resulting uncertainties into any PDF-dependent observable, preserving their simple probabilistic interpretation. In our analysis we include a broad set of deep inelastic scattering data from HERA, BCDMS and NMC experiments along with the Drell-Yan, 𝑊 and 𝑍 boson data from LHC and Tevatron experiments, which combined with theoretical calculations at next-to-next-to-leading order in QCD allow for realistic determination of PDFs. The main focus of this analysis is to explore alternative methods for PDF uncertainty estimation that are more firmly grounded in statistical principles. We show that the flexibility of the Bayes framework, allowing one, e.g., to account for non-Gaussianity or inconsistencies of datasets, is crucial to extract realistic uncertainties when such assumptions are not fulfilled. We also demonstrate that MCMC allows one to determine the Δ⁢𝜒 2 value corresponding to a given confidence level in the sample, which can, in turn, be used as a statistically well-founded tolerance criterion used in the Hessian method, thus addressing one of its main long-standing drawbacks.

Risse, Peter Clemens [Universität Münster (Germany

Surrogate model evaluation and building energy benchmarking for commercial buildings

Building energy consumption benchmarking involves challenges associated with various energy patterns for different building types; heating, ventilating, and air-conditioning (HVAC) system types; and climates. Given significant variation in energy use patterns, accurate prediction of long-term energy use using surrogate models remains challenging. Multiple linear regression (MLR) is commonly used for building energy benchmarking because of its simple structure; however, it lacks accuracy compared to other black-box models. Although many studies have compared surrogate models and offer guidance on model selection based on metrics, they do not provide detailed analysis on improving the surrogate model accuracy. In this paper, we implement a surrogate model using polynomial ridge regression (i.e., MLR with interaction terms combined with ridge regularization) for small office and retail strip mall buildings across six HVAC system types and all climate zones, for electricity and natural gas in baseline and proposed scenarios. A simulation workflow is developed using OpenStudio TM /EnergyPlus TM to generate simulation data using measures over a wide range of efficiency inputs. Enhancements based on statistical insights are used for improving the model accuracy using filters, input transformations, and change points. Surrogate models achieved average coefficient of variation of the root mean squared error (CVRMSE) values of 2.17, 1.06, 2.05, and 3.26 for proposed electricity, proposed natural gas, baseline electricity, and baseline natural gas, respectively, with enhancements reducing CVRMSE by an average of 14.9% across all combinations. We provide model interpretation via Shapley additive explanations to determine which input variables most influence energy consumption and provide supportive arguments for enhancements.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI