Search NASA⌕ Search

SEARCH · Search NASA

Results for “Benchmark data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 523 records · Page 29

An entropy-based debiasing approach to quantifying experimental coverage for novel applications of interest in the nuclear community

This manuscript proposes a novel information-theoretic approach to the quantification of experimental relevance, i.e., coverage, to achieve optimal data assimilation results for nuclear engineering applications. Specifically, this work posits the need for a new metric, called coverage (q C ) of an application’s quantity of interest, i.e., eigenvalue or power peaking for an advanced reactor concept, defined herein as the theoretically maximum achievable reduction in the quantity’s uncertainty given measurements from a pool of experiments in a manner that is independent of the data assimilation procedure employed. Currently, reduction in a quantity’s uncertainty is strongly biased by the underlying assumptions of the assimilation procedure to account for the under-determined nature of such problems and the similarity criterion employed to identify relevant experiments. To address this challenge, this work has developed a coverage metric, q C , based on mutual information, which establishes a new conceptual framework for assessing coverage, one that is independent of the model parameters and responses degree of variations in both the experimental and application domains, i.e., linear vs non-linear, and their prior uncertainty distributions, i.e., Gaussian vs. non-Gaussian. The q C is an entropic measure capable of addressing coverage for general nonlinear problems with non-Gaussian uncertainties and inclusive of the measurement uncertainties from multiple experiments. Numerical experiments from manufactured analytical problems as well as a set of benchmarks from the ICSBEP handbook are employed to demonstrate its theoretical and practical performance as compared to the c k -based experiment selection methodology, commonly employed in the neutronic community. The manuscript then employs other well-known adaptations to existing data assimilation methodologies for nonlinear and non-Gaussian problems capable of achieving the coverage posited by q C .

Bayesian data assimilation↗

Multiphysics Modeling of Microreactors with NEAMS codes, and Validation Based on KRUSTY Reactivity Insertion

The NEAMS Multiphysics Applications team continues to assess code usability and functionality for microreactor design and safety analyses, while demonstrating that NEAMS tools capture both steady-state and transient behavior across distinct microreactor concepts. In FY2025, the team advanced full-core, high-fidelity, multiphysics models that solve more complex problems and strengthen verification/validation for several microreactor systems: heat-pipe microreactor (HPMR), gas-cooled microreactor (GCMR), and the KRUSTY experiment. These models employ the MOOSE MultiApp/Transfers architecture with Griffin for neutronics, BISON for heat conduction/thermomechanics, Sockeye for heat pipes, SAM/THM for coolant channels and loops, and SWIFT for hydride behavior, with meshes generated via the MOOSE Reactor Module. The graphite models available in the Grizzly code were also investigated for future analyses. For the HPMR, a Na-HPMR variant was constructed to align with recently validated heat-pipe experiments and Sockeye’s LCVF capability, enabling mechanistic heat-pipe transients and startup modeling. The Na-HPMR will serve as the primary model for HPMR investigations in upcoming tasks. The load-following and single heat-pipe failure scenarios (Griffin/BISON/Sockeye), which were previously modeled for the K-HPMR, were replicated for the Na-HPMR, showing strong negative temperature feedback and highly localized thermal effects, respectively, while the startup case captured vapor-front progression and heat-removal activation. Solid mechanics was added to the previously built K-HPMR full-core model in BISON, showing minimal impact on steady-state reactivity yet enabling stress-field predictions that prepare the path for full-core TRISO performance analyses. For the GCMR, automated steady-state and four transient scenarios were executed using Griffin/BISON/SAM/SWIFT. Results confirm robust inherent safety: power collapses promptly in loss-of-cooling events, the inlet-temperature drop settles to a new equilibrium, and a single-channel blockage yields only a ~30 K local fuel-temperature rise with <0.4% power decrease. SWIFT-predicted hydrogen redistribution affects reactivity during both steady-state and transient conditions, underscoring its importance. A Brayton-cycle balance of plant (BOP) model in SAM/THM demonstrated stable startup behavior, and xenon-driven reactivity during load following was analyzed. To improve TRISO-compact temperature fidelity, a fast multiscale Heat Source Decomposition (HSD) treatment was implemented. Against heterogeneous benchmarks, HSD reduces underprediction of kernel temperatures and lowers predicted peak powers in reactivity-insertion transients compared to previous homogenized models. KRUSTY warm-critical validation progressed from FY2024 baselines: the 15Ȼ insertion shows excellent agreement in peak power (~2% high) and temperature trends, and the 30Ȼ case was automated via a feedback controller that maintained power near 3 kW for ~150 s with close agreement to data. The successful modeling of the warm critical tests has laid a strong foundation for simulating more complex nuclear system tests in the years ahead. Throughout FY2025, developer feedback was provided (e.g., MOOSE batch mesh generation, distributed pre-split meshes, Griffin sweeper on displaced meshes), several new models were contributed to the Virtual Test Bed, and an OECD-NEA WPRS multiphysics benchmark based on the HPMR was initiated to enable broader cross-comparison and best-practice development with the nuclear community at large.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

A Practical Probabilistic Benchmark for AI Weather Models

Since the weather is chaotic, it is necessary to forecast an ensemble of future states. Recently, multiple AI weather models have emerged claiming breakthroughs in deterministic skill. Unfortunately, it is hard to fairly compare ensembles of AI forecasts because variations in ensembling methodology become confounding and the baseline data volume is immense. We address this by scoring lagged initial condition ensembles—whereby an ensemble can be constructed from a library of deterministic hindcasts. This allows the first parameter‐free intercomparison of leading AI weather models' probabilistic skill against an operational baseline. Lagged ensembles of the two leading AI weather models, GraphCast and Pangu, perform similarly even though the former outperforms the latter in deterministic scoring. These results are elaborated upon by sensitivity tests showing that commonly used multiple time‐step loss functions damage ensemble calibration.

54 ENVIRONMENTAL SCIENCES↗

Assessment of RELAP5-3D Code with Molten Salt Heat Transfer Experiments

The accurate thermal-hydraulic assessment of thermal storage systems is crucial for addressing the integrity and performance of the storage system design. This is particularly important for thermal storage systems using molten salts as a heat transport and storage medium, as the chemical corrosion and erosion by the high-temperature molten salts can cause serious damage to the structural components of the storage system. To understand these effects for TerraPower’s Natrium® Demonstration Reactor design, the system thermal-hydraulic analysis code, RELAP5-3D, is utilized to analyze the thermal storage system. This study evaluates two heat transfer correlations implemented in the RELAP5-3D code—the Dittus-Boelter and Gnielinski correlations—using selected benchmark cases from two molten salt heat transfer experiments conducted at Xi’an Jiao Tong University and the German Aerospace Center. The Nusselt numbers calculated using the Gnielinski correlation agree with the experimental data within ±5% of the relative deviations at 300°C and 400°C. With the increasing Reynolds numbers, the Dittus-Boelter correlation underestimates the Nusselt numbers by up to 22% compared to the Gnielinski correlation. Comparison of the RELAP5-3D calculation results with experimental data at a high temperature of 550°C revealed that the thermo-chemical characteristics of solar salt at high temperatures surpassing the decomposition temperatures of nitrate salts reduce the accuracy of heat transfer model of the RELAP5-3D code. Specifically, it has been observed that the temperatures of the wall surface and liquid near the wall can surpass the decomposition temperatures, even though the bulk temperature of molten salts remains significantly lower than these decomposition temperatures. In summary, this study recommends the use of the Gnielinski correlation for the heat transfer analysis of molten salt, rather than the Dittus-Boelter correlation.

21 - SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLAN↗

IRIS-MASH: Efficient Multi-device Asynchronous Multi-Stream Heterogeneous Computing

In the rapidly evolving field of high-performance computing (HPC), effectively leveraging heterogeneous devices through asynchronous task programming is paramount. This paper presents a robust asynchronous task programming model tailored for a multi-device, multi-stream execution environment that incorporates a diverse array of heterogeneous computing units, including GPUs from various vendors and other accelerators. Current state-of-the-art task programming models provide methodologies to support asynchronous task executions, but they typically handle homogeneous devices using native programming languages, while support for heterogeneous devices is limited to frameworks like OpenCL. This gap presents significant challenges in abstracting heterogeneous devices to harness their true asynchronous capabilities effectively using their native programming languages. By implementing asynchronous task execution, our model significantly boosts the performance of tiled algorithm task graphs through overlapping data transfers with computation and enabling the simultaneous execution of multiple kernels. We integrate this approach into a heterogeneous Intelligent Runtime System (IRIS) and assess its performance using a suite of tiled algorithm benchmarks from the heterogeneous math kernels library (MatRIS) based on IRIS. Experimental results demonstrate a performance improvement ranging from 1.6 × to 2 × over IRIS without asynchronous support, and a notable 22% performance enhancement compared to established runtime systems such as StarPU and PaRSEC. This approach significantly improves computation efficiency of HPC workflows and provides a solid base for future exploration and development in the area of asynchronous task programming in heterogeneous systems.

Miniskar, Narasinga Rao [ORNL] (ORCID:000000018259↗

Benchmarking DFT Accuracy in Predicting O 1s Binding Energies on Metals

X-ray photoelectron spectroscopy (XPS) is a powerful tool for probing the electronic structure and composition of materials, particularly metals and metal oxides of relevance to solar cells and catalysis. Density functional theory (DFT) is often used to support XPS peak assignments, but its reliability for predicting oxygen species is not well established. Here, we compile a large data set of experimental oxygen binding energies and evaluate corresponding DFT predictions. We find that as the binding energies of metal-bound atomic oxygen species increase, especially above ≈530 eV, there is a general decrease in the accuracy of DFTpredicted values. Thus, high-binding-energy atomic oxygen species, such as those proposed as active for selective Ag-catalyzed epoxidation, are less well represented. The chemical nature of the oxygen species also influences accuracy, with molecularly bound species more reliably captured across the entire range of energies. These findings illustrate the limitations of DFT for interpreting XPS spectra and provide a benchmark for improving computational methods.

Adsorption↗

Benchmarking machine learning strategies for phase-field problems

Abstract We present a comprehensive benchmarking framework for evaluating machine-learning approaches applied to phase-field problems. This framework focuses on four key analysis areas crucial for assessing the performance of such approaches in a systematic and structured way. Firstly, interpolation tasks are examined to identify trends in prediction accuracy and accumulation of error over simulation time. Secondly, extrapolation tasks are also evaluated according to the same metrics. Thirdly, the relationship between model performance and data requirements is investigated to understand the impact on predictions and robustness of these approaches. Finally, systematic errors are analyzed to identify specific events or inadvertent rare events triggering high errors. Quantitative metrics evaluating the local and global description of the microstructure evolution, along with other scalar metrics representative of phase-field problems, are used across these four analysis areas. This benchmarking framework provides a path to evaluate the effectiveness and limitations of machine-learning strategies applied to phase-field problems, ultimately facilitating their practical application.

36 MATERIALS SCIENCE↗

Benchmarking the Suitability of Novec$^{\mathrm{TM}}$ 4710 for Application in Flux Compression Generators

Here, an experimental study evaluated the feasibility of replacing traditional insulating gases such as SF 6 with C 4 F 7 N (3M TM , Novec 4710) in flux compression generator (FCG) applications. Currently available data indicate that Novec 4710 could offer certain performance benefits over SF 6 . However, the available literature is focused on low frequency (50–60 Hz) and dc at static pressures. To evaluate the performance of Novec 4710 under the pulsed dynamic pressure and temperature conditions found in an FCG, we report a performance comparison between three sets of identical FCGs using air, SF 6 , and Novec 4710 as the insulating gas. The generators used in this study had a single stage, directly seeded design with an armature diameter of 25 mm and a stator diameter of 46 mm. To highlight the performance of the different gases rather than any wire insulation, the stator was constructed with uninsulated wire. Furthermore, the generators were seeded aggressively, making the performance difference between the different gases more apparent. The performance was monitored with a pair of differential Rogowski coils that captured the generators’ di / dt while also using high-speed videography to capture possible gaseous breakdown signatures. The data gathered during this study indicate that Novec 4710 performs at least as well as SF 6 in FCG applications, if not significantly better.

42 ENGINEERING↗

Combination of searches for heavy spin-1 resonances using 139 fb -1 of proton-proton collision data at $\sqrt{s}$ = 13 TeV with the ATLAS detector

A combination of searches for new heavy spin-1 resonances decaying into different pairings of W, Z, or Higgs bosons, as well as directly into leptons or quarks, is presented. The data sample used corresponds to 139 fb -1 of proton-proton collisions at $\sqrt{s}$ = 13 TeV collected during 2015–2018 with the ATLAS detector at the CERN Large Hadron Collider. Analyses selecting quark pairs (qq, bb, $t\overline{t}$, and tb) or third-generation leptons (τν and ττ) are included in this kind of combination for the first time. A simplified model predicting a spin-1 heavy vector-boson triplet is used. Cross-section limits are set at the 95% confidence level and are compared with predictions for the benchmark model. These limits are also expressed in terms of constraints on couplings of the heavy vector-boson triplet to quarks, leptons, and the Higgs boson. The complementarity of the various analyses increases the sensitivity to new physics, and the resulting constraints are stronger than those from any individual analysis considered. The data exclude a heavy vector-boson triplet with mass below 5.8 TeV in a weakly coupled scenario, below 4.4 TeV in a strongly coupled scenario, and up to 1.5 TeV in the case of production via vector-boson fusion.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Search for Beyond the Standard Model physics with anomaly detection in multilepton final states in pp collisions at s=13TeV with the ATLAS detector

A model-agnostic search for Beyond the Standard Model physics is presented, targeting final states with at least four light leptons (electrons or muons). The search regions are separated by event topology and unsupervised machine learning is used to identify anomalous events in the full 140 fb-1$$^{-1}$$ of proton–proton collision data collected with the ATLAS detector during Run 2. No significant excess above the Standard Model background expectation is observed. Model-agnostic limits are presented in each topology, along with limits on several benchmark models including vector-like leptons, wino-like charginos and neutralinos, or smuons. Limits are set on the flavourful vector-like lepton model for the first time.

Aad, G↗

Toward a persistent event-streaming system for high-performance computing applications

High-performance computing (HPC) applications have traditionally relied on parallel file systems and file transfer services to manage data movement and storage. Alternative approaches have been proposed that use direct communications between application components, trading persistence and fault tolerance for speed. Event-driven architectures, as popularized in enterprise contexts, present a compelling middle ground, avoiding the performance cost and API constraints of parallel file systems while retaining persistence and offering impedance matching between application components. However, adapting streaming frameworks to HPC workloads requires addressing challenges unique to HPC systems. This paper investigates the potential for a streaming framework designed for HPC infrastructures and use cases. We introduce Mofka, a persistent event-streaming framework designed specifically for HPC environments. Mofka combines the capabilities of a traditional streaming service with optimizations tailored to the HPC context, such as support for massively multicore nodes, efficient scaling for large producer-consumer workflows, RDMA-enabled high-performance network communications, specialized network fabrics with multiple links per node, and efficient handling of large scientific data payloads. Built using the Mochi suite of HPC data service components, Mofka provides a lightweight, modular, and high-performance solution for persistent streaming in HPC systems. We present the architecture of Mofka and evaluate its performance against Kafka and Redpanda using benchmarks on diverse platforms, including Argonne's Polaris and Oak Ridge's Frontier supercomputers, showing up to 8× improvement in throughput in some scenarios. We then demonstrate its utility in several real-world applications: a tomographic reconstruction pipeline, a workflow for the discovery of metal-organic frameworks for carbon capture, and the instrumentation of Dask workflows for provenance tracking and performance analysis.

HPC↗

RELAP5-3D HTGR Validation Work at Idaho National Laboratory

Prismatic block-type high-temperature gas-cooled reactors (HTGRs) were built in the United States decades ago, and now advanced reactor vendors are seeking to deploy them again for a variety of applications. Deploying these reactors requires modelling and simulation tools that have been validated against conditions representative of the HTGR application. Idaho National Laboratory (INL) is leading the execution an of HTGR thermal hydraulics benchmark to accelerate the validation of thermal hydraulics modelling and simulation tools for these applications. That benchmark is based on a facility called the High Temperature Test Facility (HTTF). This work provides an overview of work conducted at INL over the last 2 years to validate RELAP5-3D against data from HTTF. This presentation shows results from multiple RELAP5-3D models and an HTTF experiment to assess the impact of certain modelling assumptions on results. The contents of this talk sit on the cutting edge of RELAP5-3D validation for HTGR analysis.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Aggregation Methods for Quantifying PTM and Structural Changes in Bottom-Up Proteomics

Bottom-up proteomic workflows rely on sequential preprocessing steps, commonly including peptide-to-protein aggregation (“roll-up”), to enhance data reliability and interpretability. While roll-up is effective for protein-centered analyses, it may be suboptimal for applications focused on post-translational modifications (PTMs) or protein structural changes, such as limited proteolysis–mass spectrometry (LiP-MS). Here, we investigate how different roll-up strategies influence site-level quantification in PTM differential analysis. Moreover, we introduce a novel site-centric roll-up approach tailored for LiP-MS, which quantifies proteolytic fragments rather than solely tryptic peptides. We benchmark these methods through simulation studies, comparing their sensitivity and specificity in detecting structural and PTM-driven changes. We found that the median and mean roll-up methods outperform the sum method in both PTM and LiP proteomics, and site-level quantification in LiP outperforms peptide-level quantification. Our findings offer the first systematic, data-driven guidance for selecting roll-up techniques in site-level proteomic analyses, with implications for both PTM-focused and structural proteomics studies.

aggregation↗

Leveraging unlabeled SEM datasets with self-supervised learning for enhanced particle segmentation

Scanning Electron Microscopes (SEMs) are widely used in experimental science laboratories, often requiring cumbersome and repetitive user analysis. Automating SEM image analysis processes is highly desirable to address this challenge. In particle sample analysis, Machine Learning (ML) has emerged as the most effective approach for particle segmentation. However, the time-intensive process of manually annotating thousands of SEM images limits the applicability of supervised learning approaches. Self-Supervised Learning (SSL) offers a promising alternative by enabling knowledge extraction from raw, unlabeled data. This study presents a framework for evaluating SSL techniques in SEM image analysis, focusing on novel methods leveraging the ConvNeXtV2 architecture for particle detection. A dataset comprising 25,000 SEM images is curated to benchmark these proposed SSL methods. The results demonstrate that ConvNeXtV2 models, with varying parameter counts, consistently outperform other techniques in particle detection across different length scales, achieving up to a 34% reduction in relative error compared to established SSL methods. Furthermore, an ablation study explores the relationship between dataset size and SSL performance, providing actionable insights for practitioners regarding model selection and resource efficiency. This research advances the integration of SSL into autonomous analysis pipelines and supports its application in accelerating materials science discovery.

Rettenberger, Luca↗

Uncertainty quantification for molecular property predictions with graph neural architecture search

Graph Neural Networks (GNNs) have emerged as a prominent class of data-driven methods for molecular property prediction. However, a key limitation of typical GNN models is their inability to quantify uncertainties in the predictions. This capability is crucial for ensuring the trustworthy use and deployment of models in downstream tasks. To that end, we introduce AutoGNNUQ, an automated uncertainty quantification (UQ) approach for molecular property prediction. AutoGNNUQ leverages architecture search to generate an ensemble of high-performing GNNs, enabling the estimation of predictive uncertainties. Our approach employs variance decomposition to separate data (aleatoric) and model (epistemic) uncertainties, providing valuable insights for reducing them. In our computational experiments, we demonstrate that AutoGNNUQ outperforms existing UQ methods in terms of both prediction accuracy and UQ performance on multiple benchmark datasets, and generalizes well to out-of-distribution datasets. Additionally, we utilize t-SNE visualization to explore correlations between molecular features and uncertainty, offering insight for dataset improvement. AutoGNNUQ has broad applicability in domains such as drug discovery and materials science, where accurate uncertainty quantification is crucial for decision-making.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Search for Light Long-Lived Particles in $pp$ Collisions at $\sqrt{s} =$ 13 TeV Using Displaced Vertices in the ATLAS Inner Detector

A search for long-lived particles (LLPs) using 140 fb -1 of pp collision data with $\sqrt{s} =$ 13 TeV recorded by the ATLAS experiment at the LHC is presented. The search targets LLPs with masses between 5 and 55 GeV that decay hadronically in the ATLAS inner detector. Benchmark models with LLP pair production from exotic decays of the Higgs boson and models featuring long-lived axionlike particles (ALPs) are considered. No significant excess above the expected background is observed. Upper limits are placed on the branching ratio of the Higgs boson to pairs of LLPs, the cross section for ALPs produced in association with a vector boson, and, for the first time, on the branching ratio of the top quark to an ALP and a u/c quark.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Development of a Distribution Optimal Power Flow Federate for Open-Source OEDI-SI Platform

Increasing numbers of distributed generators in the electric power distribution networks require developing a control strategy to optimize solutions in real time. Linearized optimal distribution flow development has seen growth and acceptance in the distribution systems literature for efficiently modeling the \glspl{opf} for distribution systems. This paper examines the implementation and integration procedure for linearized optimal distribution flow federate to \gls{oedisi} platform. Specifically, we discuss i) the usage of the \gls{oedisi} platform, ii) obtaining a tractable solution using developed \gls{opf} federate, and iii) validation of solutions and bench-marking the \gls{oedisi} platform with developed \gls{opf} federate using OpenDSS. In brief, we demonstrate how a general linearized optimal distribution flow federate can be developed and integrated with a co-simulation environment to mimic real-world examples. The efficacy of the proposed method is demonstrated using the IEEE 123-bus test system under different scenarios to obtain a tractable solution and compare its results.

Sadnan, Rabayet↗

Differentially Private Map Matching (DPMM) v1.0

Human mobility trajectories provide valuable information for developing mobility applications, as they contain diverse and rich information about the users. User mobility data is valuable for various applications such as intelligent transportation systems (ITS), commercial business models, and disease-spread models. However, such spatio-temporal traces may pose a threat to user privacy. GPS trajectories in their raw form are not suitable for transportation studies, as they require matching locations with nearest road links — a process called map-matching. This software implements a differential privacy (DP)-based map-matching algorithm, called DPMM, that generates link-level location trajectories in a privacy-preserving manner to protect users' origin destinations (OD) and travel paths. OD privacy is achieved by injecting Planar Laplace noise to the user OD GPS points. Travel-path privacy is provided with randomized travel path construction using exponential DP mechanism. The injected noise level is selected adaptively, by considering the link density of the location and the functional category of the localized links. For path privacy, our mechanism samples waypoints and selects candidate paths between waypoints. DPMM provides privacy effectively with respect to link density instead of other trajectory samples in the database compared to other privacy mechanisms. Compared to the different baseline models our DP-based privacy model offers closer query responses to the raw data in terms of individual and aggregate trajectory-level statistics with an average at absolute deviation from the baseline for individual statistics on ϵ = 1.0. Beyond individual trajectory statistics, the DPMM outperforms the other benchmark DP-based mechanisms on different aggregate statistics with up to 8x improvement in utility.

Peisert, Sean [Lawrence Berkeley National Laborato↗