Search NASA⌕ Search

SEARCH · Search NASA

Results for “Modelling”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 487 records · Page 27

Core-edge integrated predictive studies of ST40 and NSTX plasmas with the scrape-off layer box model

The ability to model the interplay between the core and edge of tokamak plasmas is crucial to designing both the plasma operating scenario of a fusion pilot plant and the design of the tokamak itself. Scrape-off-layer (SOL) models that are tailored to integrated scenario modeling need to have fast turn-around time and minimal computational burden to enable wide parameter-space coverage for design scoping. The SOL 0-D Box model is a reduced SOL model based on global power and particle balance that captures the essential physics of SOL transport with little computational cost. The usage of the 0-D Box model in core-edge coupled simulations has been demonstrated in both interpretive and predictive modes on a variety of devices. This paper presents a sensitivity study of the 0-D Box model to the input SOL heat-flux width for an ST40 plasma. This study demonstrates that accurate prediction of this width is crucial to predicting global performance parameters of a plasma scenario, such as energy confinement time and flux consumption. We also present an extension of the Box model to 1-D to allow for parallel variation of plasma parameters along the magnetic field lines. The 1-D Box model is then compared with SOLPS-ITER simulations of an NSTX plasma. Advantages and limitations of the Box model are discussed, and future directions are outlined.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

A comparison between ShapeFit compression and Full-Modelling method with PyBird for DESI 2024 and beyond

DESI aims to provide one of the tightest constraints on cosmological parameters by analysing the clustering of more than thirty million galaxies. However, obtaining such constraints requires special care in validating the methodology and efforts to reduce the computational time required through data compression and emulation techniques. In this work, we perform a rigorous validation of the PyBird power spectrum modelling code with both a traditional emulated Full-Modelling approach and the model-independent ShapeFit compression approach. By using cubic box simulations that accurately reproduce the clustering and precision of the DESI survey, we find that the cosmological constraints from ShapeFit and Full-Modelling are consistent with each other at the ∼ 0.5σ level for the ΛCDM model. Both ShapeFit and Full-Modelling are also consistent with the true ΛCDM simulation cosmology down to a scale of k max = 0.20 hMpc -1 even after including the hexadecapole. For extended models such as the wCDM and the oCDM models, we find that including the hexadecapole can significantly improve the constraints and reduce the modelling errors with the same k max . While their discrepancies between the constraints from ShapeFit and Full-Modelling are more significant than ΛCDM, they remain consistent within 0.7σ. Lastly, we also show that the constraints on cosmological parameters with the correlation function evaluated from PyBird down to s min = 30h -1 Mpc are unbiased and consistent with the constraints from the power spectrum.

79 ASTRONOMY AND ASTROPHYSICS↗

Data-driven model validation for neutrino-nucleus cross section measurements

Neutrino-nucleus cross section measurements are needed to improve interaction modeling to meet the precision needs of neutrino experiments in efforts to measure oscillation parameters and search for physics beyond the Standard Model. We review the difficulties associated with modeling neutrino-nucleus interactions that lead to a dependence on event generators in oscillation analyses and cross section measurements alike. We then describe data-driven model validation techniques intended to address this model dependence. The method relies on utilizing various goodness-of-fit tests and the correlations between different observables and channels to probe the model for defects in the phase space relevant for the desired analysis. These techniques shed light on relevant mismodeling, allowing it to be detected before it begins to bias the cross section results. We compare more commonly used model validation methods which directly validate the model against alternative ones to these data-driven techniques and show their efficacy with fake data studies. These studies demonstrate that employing data-driven model validation in cross section measurements represents a reliable strategy to produce robust results that will stimulate the desired improvements to interaction modeling.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

A Near-Real-Time Model for Predicting Electricity Disruptions in Texas During Winter Storms

There has been an increase in extreme weather events, posing a threat to power grid systems, potentially influenced by factors such as population growth, changes in ecosystems, land cover, and land use in the service area, as well as the growth of certain vegetation types. This research seeks to develop a predictive model to mitigate potential damages caused by future winter storms. This research utilizes the Light Gradient Boosting Machine (LightGBM), incorporating the number of power outages experienced at the county level, geographic details, weather information, and lagged outage and lagged weather data. The developed models were broadly divided into two groups, with six models in each group - one group without optimization and another with optimization, totaling 12 trained models. For model optimization, Bayesian optimization was employed using Root Mean Squared Error (RMSE) as the objective function. In results, when comparing Group 2 (the optimized group) with Group 1 (the non-optimized group), it was found that optimization did not always lead to a reduction in RMSE and Mean Absolute Error (MAE). However, in terms of Mean Directional Accuracy (MDA), while all results in Group 1 were below the baseline accuracy of 0.33, all results in Group 2 exceeded 0.33, with some cases showing an increase of more than three times the baseline. The results indicated that, in the optimized model group, Population and Pressure were the most influential factors when using current weather data and geographical information. When using lagged data, lagged recorded outages and lagged Pressure emerged as the most significant factors. Among the 12 developed models, the L-1-2-O model showed the lowest RMSE and MAE, as well as the highest accuracy, with values of 390.62 households and 168.13 households, respectively. To normalize the RMSE and MAE values, each metric was divided by the average number of households among the counties in Texas. For the L-1-2-O model, the scaled RMSE was 0.88% and the scaled MAE was 0.38%. In terms of MDA, which indicates the accuracy of the prediction direction, the L-1-O model achieved the highest score of 0.41. Although this study focused on Texas, which suffered the greatest impact from the winter storms in 2021, with additional validation, the methodology used in this research could be applied to other regions.

Lee, Jangjae [Texas A & M Univ., College Station, ↗

Pretraining Billion-Scale Geospatial Foundational Models on Frontier

As AI workloads increase in scope, generalization capability becomes challenging for small task-specific models and their demand for large amounts of labeled training samples increases. On the contrary, Foundation Models (FMs) are trained with internet-scale unlabeled data via self-supervised learning and have been shown to adapt to various tasks with minimal fine-tuning. Although large FMs have demonstrated significant impact in natural language processing and computer vision, efforts toward FMs for geospatial applications have been restricted to smaller size models, as pretraining larger models requires very large computing resources equipped with state-of-the-art hardware accelerators. Current satellite constellations collect 100+TBs of data a day, resulting in images that are billions of pixels and multimodal in nature. Such geospatial data poses unique challenges opening up new opportunities to develop FMs. We investigate billion scale FMs and HPC training profiles for geospatial applications by pretraining on publicly available data. We studied from end-to-end the performance and impact in the solution by scaling the model size. Our larger 3B parameter size model achieves up to 30% improvement in top1 scene classification accuracy when comparing a 100M parameter model. Moreover, we detail performance experiments on the Frontier supercomputer, America's first exascale system, where we study different model and data parallel approaches using PyTorch's Fully Sharded Data Parallel library. Specifically, we study variants of the Vision Transformer architecture (ViT), conducting performance analysis for ViT models with size up to 15B parameters. By discussing throughput and performance bottlenecks under different parallelism configurations, we offer insights on how to leverage such leadership-class HPC resources when developing large models for geospatial imagery applications.

Tsaris, Aristeidis (aris)↗

Boron Coordination in Multicomponent Glasses: Analytical Models and Machine Learning With Uncertainty

Borosilicate glasses are extensively used in a variety of applications from kitchenware to nuclear waste immobilization due to the strong network formed by the Si-O-B bond that makes it resistant to chemical corrosion and gives it a low thermal expansion. Boron, however, exists in both trigonal BO3 and tetrahedral BO4 bonds in glass systems, which impacts the chemical durability and thermal resistance of the glass, amongst other properties. Boron coordination (N4), or the ratio of the amount of BO4 to BO3 within a glass, may aid in predicting these properties but is difficult to derive without experimental data due to the complexity of impacts from varied glass compositions and processing factors. For this reason, compositional models have been developed to predict boron coordination, but the models typically include a limited number of glass components. To help fill this gap in the models, in this work, a diverse multicomponent glass dataset of 809 glasses is compiled from a literature search, and then a number of analytical and machine learning (ML) models are trained on the dataset. Previously developed modified Bernstein and modified Du Stebbins analytical models were fitted to update parameters with the new dataset. Then, partially Bayesian neural networks, Gaussian process regressor, and heteroskedastic deterministic neural networks were evaluated. The ML models examined all have different strategies to overcome the potential for overfitting as a result of a limited training dataset, and return results that account for model uncertainty, which can be valuable for understanding model reliability. For the first time, cooling rate is introduced as an input parameter for ML models, showing consistent improvements in performance and solidifying the importance of including parameters outside of composition alone for N4 prediction. The machine learning models examined here show promise in accurate predictions of boron coordination in borosilicate glasses, all achieving R2 values of 0.91.

boron coordination↗

Coupled Multiphysics Modeling of Lithium-Ion Batteries for Automotive Crashworthiness Applications

Considerable advances have been made in battery safety models, but achieving predictive accuracy across a wide range of conditions continues to be challenging. Interactions between dynamically evolving mechanical, electrical, and thermal state variables make model prediction difficult during mechanical abuse scenarios. In this study, we develop a physics-based modeling approach that allows for choosing between different mechanical and electrochemical models depending on the required level of analysis. We demonstrate the use of this approach to connect cell-level abuse response to electrode-level and particle-level transport phenomena. A pseudo-two-dimensional model and simplified single-particle models are calibrated to electrical-thermal cycling data and applied to mechanically induced short-circuit scenarios to understand how the choice of electrochemical model affects the model prediction under abuse scenarios. These models are implemented using user-defined subroutines on ls-dyna finite element software and can be coupled with existing automotive crash safety models.

analysis and design of components↗

Meta Biome: a multiscale model integrating agent-based and metabolic networks to reveal spatial regulation in gut mucosal microbial communities

ABSTRACT Mucosal microbial communities (MMCs) are complex ecosystems near the mucosal layers of the gut essential for maintaining health and modulating disease states. Despite advances in high-throughput omics technologies, current methodologies struggle to capture the dynamic metabolic interactions and spatiotemporal variations within MMCs. In this work, we presentMetaBiome, a multiscale model integrating agent-based modeling (ABM), finite volume methods, and constraint-based models to explore the metabolic interactions within these communities. Integrating ABM allows for the detailed representation of individual microbial agents each governed by rules that dictate cell growth, division, and interactions with their surroundings. Through a layered approach—encompassing microenvironmental conditions, agent information, and metabolic pathways—we simulated different communities to showcase the potential of the model. Using ourin-silicoplatform, we explored the dynamics and spatiotemporal patterns of MMCs in the proximal small intestine and the cecum, simulating the physiological conditions of the two gut regions. Our findings revealed how specific microbes adapt their metabolic processes based on substrate availability and local environmental conditions, shedding light on spatial metabolite regulation and informing targeted therapies for localized gut diseases.MetaBiome provides a detailed representation of microbial agents and their interactions, surpassing the limitations of traditional grid-based systems. This work marks a significant advancement in microbial ecology, as it offers new insights into predicting and analyzing microbial communities. IMPORTANCE Our study presents a novel multiscale model that combines agent-based modeling, finite volume methods, and genome-scale metabolic models to simulate the complex dynamics of mucosal microbial communities in the gut. This integrated approach allows us to capture spatial and temporal variations in microbial interactions and metabolism that are difficult to study experimentally. Key findings from our model include the following: (i) prediction of metabolic cross-feeding and spatial organization in multi-species communities, (ii) insights into how oxygen gradients and nutrient availability shape community composition in different gut regions, and (iii) identification of spatiallyregulated metabolic pathways and enzymes inE. coli. We believe this work represents a significant advance in computational modeling of microbial communities and provides new insights into the spatial regulation of gut microbiome metabolism. The multiscale modeling approach we have developed could be broadly applicable for studying other complex microbial ecosystems.

Microbiology↗

Improving the Transportability of a Deep Learning Denoising Model Using Transfer Learning Techniques

The adoption of machine learning techniques in the seismology community has led to great performance improvements in several areas, including signal processing. Specifically, the development of deep learning–based seismic waveform denoising models has the potential to yield improvements in signal detection capabilities for networks operating in particularly noisy environments. Recent advancements in the design of these deep learning denoising models have included the incorporation of continuous and discrete wavelet transform functions into the network architecture to improve the learning capabilities and efficiency of said models. These wavelet transform–based seismic denoising models have shown improved denoising capabilities in regions where there is good agreement between the data features present in the training and evaluation datasets. However, questions remain about the overall transportability of these models to other monitoring regions. Here, in this study, we will determine the baseline transportability of a newly developed multilevel wavelet‐transform convolutional neural network (MWCNN) seismic denoising model. We accomplish this by taking a version of the MWCNN denoising model trained on data collected from the Utah region and evaluating its denoising performance on datasets collected from the neighboring Nevada region, which differ with regard to monitoring sensor types and event histories. We find that there is a notable variability in denoising performance related to the degree of similarity between the initial and new target datasets. The most notable difference in denoising performance is the ability of the denoising model to preserve accurate amplitude information associated with the signal energy present in the waveform data. Finally, we evaluate the ability of transfer learning techniques to improve the transportability of the MWCNN denoising model. We find that although there is still a performance gap present in the denoising results of the MWCNN model, transfer learning did yield improved results.

Quinones, Louis [Sandia National Laboratories (SNL↗

New Virtual Test Bed Capabilities: Virtual DOME Model and New Updates to Repository

The Department of Energy (DOE) Office of Nuclear Energy National Reactor Innovation Center accelerates the deployment of novel reactor concepts by establishing both physical and virtual spaces for building and testing various components, systems, and complete pilot plants. The Virtual Test Bed represents the virtual arm of the National Reactor Innovation Center and is a joint effort with the DOE Nuclear Energy Advanced Modeling and Simulation Program. The Virtual Test Bed mission is to accelerate the deployment of advanced reactors by facilitating the adoption of cutting-edge DOE advanced modeling and simulation tools to design, evaluate, and license reactors. This is primarily achieved by storing example challenge problems in an externally available repository and by developing models to fill the M&S gaps needed for potential demonstrators. Activities conducted this fiscal year focused on developing of a Demonstration of Microreactor Experiments shield model to help accelerate the confirmatory analysis required for the reactor demonstration. This model and workflow will allow developers to leverage advanced modeling and simulation tools to ensure their reactor demonstration concept will meet dose requirements and that the surrounding shield will stay within concrete temperature limits during steady-state and transient operation conditions. An initial model has been developed to evaluate the temperature distribution in the concrete shield during steady-state operation, including neutron and gamma heating effects. Various modeling strategies have been examined to understand their applicability and limitations with different reactor designs to make the workflow as reactor-agnostic as possible and computationally effective to maximize its usability. In addition to describing the Demonstration of Microreactor Experiments shield model and associated results, this report summarizes other accomplishments regarding repository maintenance and improvement and new external models hosted on the repository.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

A Semi-Detailed Pyrolytic Gas-Phase Kinetic Model for the Volatiles of Polyethylene Thermal Degradation

This work presents a semi-detailed kinetic model to address the pyrolytic gas-phase reactivity of volatiles formed during thermal degradation of polyethylene (PE). The model builds on a validated multi-step condensed-phase model and employs validated lumping approaches. Short-chain compounds are modelled with high detail, while long-chain ones are described by surrogate species representative of diesel-cuts (NC16H32) and waxes (NC30H60). The reactivity of short chains is described through the comprehensive CRECK kinetic model, updated to align C5-C7 olefins based on recent literature experimental data. Due to the lack of experimental data for longer olefins, their reactivity is modeled by analogy to the shorter ones, ensuring an asymptotic behavior with increasing carbon numbers. The semi-detailed model is validated through experimental data on PE pyrolysis, assuming an instantaneous mixing of the inert inlet flow with released volatiles, followed by a segregated plug-flow behavior. Validation across different reactor setups confirms the model’s capability to predict detailed product distributions. Despite minor discrepancies, the proposed model effectively captures experimental trends. Further work will address modelling the reactivity in oxygen-containing environments.

kinetics↗

ASME Code change proposal to implement new universal high temperature constitutive models for Section III, Division 5

This report completes work on a universal high temperature constitutive model suitable for use with the ASME Boiler & Pressure Vessel Code Section III, Division 5 rules for the design by inelastic analysis of Class A nuclear reactor components. The goals of this work are to provide a simple model form that adequately captures the high temperature response of materials and can be applied to any future Code material. Additionally, the report describes an automated process for calibrating a model against test data. The idea is to simplify the effort required to generate a constitutive model for an arbitrary material, provided test data is available. This will accelerate the process of qualifying new Code materials in the future. In addition, the report provides calibrated models and detailed validation comparisons to test data for five currently-qualified or soon-to-be qualified materials: 316H, Grade 91, Alloy 800H, Alloy 617, and Alloy 709. The report surveys the available data for the remaining two ASME Class Materials --- 2.25Cr-1Mo and 304H --- concluding that there is enough data data to generate a model for 2.25Cr-1Mo steel provided some additional sources of non-public data can be included in the test database, but that a dedicated cyclic test campaign would be needed for 304H. Supplemental material includes the full text of an ASME Code change proposal to incorporate the models for the four currently-qualified Class A material, detailed validation comparisons to test data for the five material models, and input files for reference implementations of the constitutive models in the NEML and NEML2 modeling frameworks.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Electromagnetic Transient Modeling of Large Data Centers for Grid-Level Studies

The magnitude and complexity of electricity usage patterns from large data centers are having significant impacts on the operation and dynamics of the power grid; grid operators and planners require a range of specialized data center models to properly evaluate these impacts and specify technical solutions as needed. Towards addressing this need, Pacific Northwest National Laboratory (PNNL) has developed a library of electromagnetic transient (EMT) models for grid-level studies of data centers called the data center model library (DML). This report describes how the DML was created and how it may properly be used. The models present in the DML are generic models; subject matter expertise and additional technical data are needed to modify these models before they can represent any real data center. However, they will significantly reduce the level of effort required to develop site-specific models and can serve as a common starting point to guide industry towards a more refined consensus. Most of the models within DML are dedicated to representing the power electronics interfaces commonly used in modern data centers, such as double-conversion uninterruptible power supplies and single-phase power factor correction converters. These models are intended for use in grid-level studies and are a simplified aggregation of many small components. That said, background material on the physical and electrical design of large data centers is provided as companion material so that users can be aware of many of the details which have been omitted or streamlined as a matter of practical necessity. Additionally, guidance on the application of EMT analysis for data center interconnection studies is provided, which aids users in identifying when the DML is necessary and what sort of additional model development may be necessary for conducting real-world studies.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Modeling Heat Pipe Startup And Noncondensable Gases In Sockeye

A one-dimensional gas mixture flow model was developed and implemented in the heat pipe code Sockeye to model the effects of non-condensable gases. Additionally a startup model based on the dusty gas model was implemented to model the transition from rarefied gas dynamics to continuum flow, which occurs during the frozen startup of high-temperature heat pipes. Multiple startup models and non-condensable gas models were tested against experimental data for sodium heat pipes, showing excellent agreement. Additionally, the newly developed gas mixture model for modeling non-condensable gas is further tested with a theoretical case study with arbitrary heating configurations. Finally, several recommendations and conclusions are made from the studies in this work to guide future heat pipe modeling efforts.

97 - MATHEMATICS AND COMPUTING↗

Electromagnetic Transient Modeling of Large Data Centers for Grid-Level Studies: Beta Release

The magnitude and complexity of electricity usage patterns from large data centers are having significant impacts on the operation and dynamics of the power grid; grid operators and planners require a range of specialized data center models to properly evaluate these impacts and specify technical solutions as needed. Towards addressing this need, Pacific Northwest National Laboratory (PNNL) has developed a library of electromagnetic transient (EMT) models for grid-level studies of data centers called the data center model library (DML). This report describes how the DML was created and how it may properly be used. This report details the DML’s beta release, completed in July 2026. This is a revision and expansion of the alpha release, which was made available in January 2026 The models present in the DML are generic models; subject matter expertise and additional technical data are needed to modify these models before they can represent any real data center. However, they will significantly reduce the level of effort required to develop site-specific models and can serve as a common starting point to guide industry towards a more refined consensus. Most of the models within DML are dedicated to representing the power electronics interfaces commonly used in modern data centers, such as double-conversion uninterruptible power supplies and single-phase power factor correction converters. These models are intended for use in grid-level studies and are a simplified aggregation of many small components. That said, background material on the physical and electrical design of large data centers is provided as companion material so that users can be aware of many of the details which have been omitted or streamlined as a matter of practical necessity. Additionally, guidance on the application of EMT analysis for data center interconnection studies is provided, which aids users in identifying when the DML is necessary and what sort of additional model development may be necessary for conducting real-world studies.

electromagnetic transients↗

An Empirical Model For Intrinsic Alignments: Insights From Cosmological Simulations

We extend current models of the halo occupation distribution (HOD) to include a flexible, empirical framework for the forward modeling of the intrinsic alignment (IA) of galaxies. A primary goal of this work is to produce mock galaxy catalogs for the purpose of validating existing models and methods for the mitigation of IA in weak lensing measurements. This technique can also be used to produce new, simulation-based predictions for IA and galaxy clustering. Our model is probabilistically formulated, and rests upon the assumption that the orientations of galaxies exhibit a correlation with their host dark matter (sub)halo orientation or with their position within the halo. We examine the necessary components and phenomenology of such a model by considering the alignments between (sub)halos in a cosmological dark matter only simulation. We then validate this model for a realistic galaxy population in a set of simulations in the Illustris-TNG suite. We create an HOD mock with Illustris-like correlations using our method, constraining the associated IA model parameters, with the $\mathcal{X}$$^{2}_{dof}$ between our model’s correlations and those of Illustris matching as closely as 1.4 and 1.1 for orientation–position and orientation–orientation correlation functions, respectively. By modeling the misalignment between galaxies and their host halo, we show that the 3-dimensional two-point position and orientation correlation functions of simulated (sub)halos and galaxies can be accurately reproduced from quasi-linear scales down to 0.1 $\mathcal{h}$ –1 Mpc. We also find evidence for environmental influence on IA within a halo. Our publicly-available software provides a key component enabling efficient determination of Bayesian posteriors on IA model parameters using observational measurements of galaxy-orientation correlation functions in the highly nonlinear regime.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Comparison of Machine Learning-Based Predictive Models of the Nutrient Loads Delivered from the Mississippi/Atchafalaya River Basin to the Gulf of Mexico

Predicting nutrient loads is essential to understanding and managing one of the environmental issues faced by the northern Gulf of Mexico hypoxic zone, which poses a severe threat to the Gulf’s healthy ecosystem and economy. The development of hypoxia in the Gulf of Mexico is strongly associated with the eutrophication process initiated by excessive nutrient loads. Due to the complexities in the excessive nutrient loads to the Gulf of Mexico, it is challenging to understand and predict the underlying temporal variation of nutrient loads. The study was aimed at identifying an optimal predictive machine learning model to capture and predict nonlinear behavior of the nutrient loads delivered from the Mississippi/Atchafalaya River Basin (MARB) to the Gulf of Mexico. For this purpose, monthly nutrient loads (N and P) in tons were collected from US Geological Survey (USGS) monitoring station 07373420 from 1980 to 2020. Machine learning models—including autoregressive integrated moving average (ARIMA), gaussian process regression (GPR), single-layer multilayer perceptron (MLP), and a long short-term memory (LSTM) with the single hidden layer—were developed to predict the monthly nutrient loads, and model performances were evaluated by standard assessment metrics—Root Mean Square Error (RMSE) and Correlation Coefficient (R). The residuals of predictive models were examined by the Durbin–Watson statistic. The results showed that MLP and LSTM persistently achieved better accuracy in predicting monthly TN and TP loads compared to GPR and ARIMA. In addition, GPR models achieved slightly better test RMSE score than ARIMA models while their correlation coefficients are much lower than ARIMA models. Moreover, MLP performed slightly better than LSTM in predicting monthly TP loads while LSTM slightly outperformed for TN loads. Furthermore, it was found that the optimizer and number of inputs didn’t show effects on the LSTM performance while they exhibited impacts on MLP outcomes. This study explores the capability of machine learning models to accurately predict nonlinearly fluctuating nutrient loads delivered to the Gulf of Mexico. Further efforts focus on improving the accuracy of forecasting using hybrid models which combine several machine learning models with superior predictive performance for nutrient fluxes throughout the MARB.

54 ENVIRONMENTAL SCIENCES↗

Data-driven model validation for neutrino-nucleus cross section measurements

Neutrino-nucleus cross section measurements are needed to improve interaction modeling to meet the precision needs of neutrino experiments in efforts to measure oscillation parameters and search for physics beyond the Standard Model. We review the difficulties associated with modeling neutrino-nucleus interactions that lead to a dependence on event generators in oscillation analyses and cross section measurements alike. We then describe data-driven model validation techniques intended to address this model dependence. The method relies on utilizing various goodness-of-fit tests and the correlations between different observables and channels to probe the model for defects in the phase space relevant for the desired analysis. These techniques shed light on relevant mismodeling, allowing it to be detected before it begins to bias the cross section results. We compare more commonly used model validation methods which directly validate the model against alternative ones to these data-driven techniques and show their efficacy with fake data studies. These studies demonstrate that employing data-driven model validation in cross section measurements represents a reliable strategy to produce robust results that will stimulate the desired improvements to interaction modeling.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗