Search NASA⌕ Search

SEARCH · Search NASA

Results for “Benchmark data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 145 records · Page 8

High-resolution spectroscopy of barium monofluoride: Odd isotopologues, hyperfine structure, and isotope shifts

Barium monofluoride (BaF) is a promising molecular species for precision tests of fundamental symmetries and interactions. We present a combined theoretical and experimental study of BaF spectra and isotope shifts, focusing in particular on the poorly understood odd isotopologues 137 BaF and 135 BaF. By comparing state-of-the-art ab initio calculations with high-resolution fluorescence and absorption spectroscopy data, we provide a benchmark for electronic structure theory and disentangle the hyperfine and rovibrational spectra of the five most abundant isotopologues, from 138 BaF to 134 BaF. The comprehensive knowledge gained enables a King plot analysis of the isotope shifts that reveals the odd-even staggering of the barium nuclear charge radii. It also paths the way for improved laser cooling of rare BaF isotopologues and crucially supports future measurements of nuclear anapole and Schiff moments.

Kogel, Felix [Univ. of Stuttgart (Germany)] (ORCID↗

Isolating Unisolated Upsilons with Anomaly Detection in CMS Open Data

We present the first study of anti-isolated Upsilon decays to two muons (ϒ→𝜇⁺⁢𝜇⁻) in proton-proton collisions at the Large Hadron Collider. Using a machine learning (ML)-based anomaly detection strategy, we “rediscover” the ϒ in 13 TeV CMS Open Data from 2016, despite overwhelming anti-isolated backgrounds. We elevate the signal significance to 6.4⁢𝜎 using these methods, starting from 1.6⁢𝜎 using the dimuon mass spectrum alone. Moreover, we demonstrate improved sensitivity from using an ML-based estimate of the multifeature likelihood compared to traditional “cut-and-count” methods. This is the first ever detection of anti-isolated Upsilons, which can be useful in the study of heavy-flavor fragmentation in quantum chromodynamics. Our Letter demonstrates that it is possible and practical to find real signals in experimental collider data using ML-based anomaly detection, and we distill a readily accessible benchmark dataset from the CMS Open Data to facilitate future anomaly detection developments.

machine learning↗

Status of the CERBERUS Evaluation for the International Criticality Safety Benchmark Evaluation Project (ICSBEP) Handbook

Modeling & Simulation (M&S) tools are used to analyze advanced reactor designs and the safety of current nuclear operations. As computers continue to improve, we are able to enhance resolution in our calculations. Therefore, the limitations of simulation capability are in the quality of data that is being used, including our ability to quantify the uncertainty and sensitivity of that data. In order to model systems of interest with increasing accuracy, the industry must improve key nuclear data measurements. The International Criticality Safety Benchmark Evaluation Project (ICSBEP) compiles and evaluates experiment data in a handbook that can be used by criticality safety engineers and others to validate computer codes and cross section libraries at nuclear facilities. Both critical and subcritical experiments are included in the handbook. These experiments, along with differential measurements, can help improve the quality of nuclear data. Concerns regarding the accuracy of Cu nuclear data have been published. The large values and trend of C-E for the Zeus intermediate energy benchmark, being one of the primary examples. Furthermore, very few experiments have been designed to be sensitive to Cu (as shown in Figure 1), so an integral, critical experiment is needed to help resolve these differences.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Open database for GPD analyses

This article summarizes the main ideas behind creating an open database proposed for use in the exploration of generalized parton distributions (GPDs). This lightweight database is well suited for GPD phenomenology and is designed to store both experimental and lattice-QCD data. It can also aid in benchmarking GPD-related developments, such as GPD models. The database utilizes a new data format based on the YAML serialization language, enabling the storage of essential information for modern analyses, such as replica values. It includes interfaces for both Python and C++, allowing straightforward integration with analysis codes.

Burkert, V. D. [Thomas Jefferson National Accelera↗

CoderData

Benchmark dataset that harmonizes drug response data across thousands of patient samples and cancer model systems enabling the training and benchmarking of AI/ML algorithms at scale.

Gosline, Sara↗

Thermodynamic Modeling of Intrinsic Defects in MnBi₂Te₄

This repository contains the computational data supporting the manuscript titled “The critical role of intrinsic defects and many-body interactions on the stability of MnBi₂Te₄.” It includes: 1. DFT data generated using VASP, used for training and benchmarking electronic structure models. 2. Quantum Monte Carlo (QMC) data produced with QMCPACK, used to apply many-body corrections and validate the electronic and magnetic properties of MnBi₂Te₄. 3. Relevant scripts used to run, analyze, and process the calculations, enabling reproducibility and transparency of the workflows.

36 MATERIALS SCIENCE↗

Benchmark of the Chlorine Worth Study Experiments in Support of Chlorine Nuclear Data Validation for Nuclear Criticality Safety

The Chlorine Worth Study (CWS) was a critical experiment to address an urgent need for thermal chlorine nuclear data validation in plutonium systems. This urgent need is tied directly to plutonium recycle and recovery operations in the plutonium facility at Los Alamos National Laboratory, where exceptionally conservative criticality safety limits are used because no credit is taken for the neutron capture by chlorine. The experiment used weapons-grade plutonium metal plates clad in stainless steel, known as the PANN (plutonium aluminum no nickel) ZPPR (zero power physics reactor) plates. The plutonium was reflected and moderated by high-density polyethylene and included combinations of polyvinyl chloride (PVC) and chlorinated polyvinyl chloride (CPVC) as absorbers. The experiment and benchmark included three configurations mimicking 30 g 239 Pu/L plutonium, 300 g 239 Pu/L plutonium, and 600 g 239 Pu/L plutonium in an aqueous chloride solution. Uncertainties in the benchmark included five broad categories: (1) criticality measurement, (2) mass and density, (3) dimensions, (4) material compositions, and (5) positioning. The largest contribution to the overall uncertainties for all three cases came from the material compositions, in particular the PVC and CPVC absorber compositions. A detailed model was created to be a near match (that is within expectations of transport code users) and a simplified model was created to minimize offset dimensions and expedite modeling for code validation. Sample calculations were completed in MCNP6.3 with ENDF/B-VIII.0 and ENDF/B-VII.1 nuclear data. For the detailed and simplified models, the average difference between the computed and experimental k eff was 951 pcm. CWS will serve as the key validation experiment for nuclear criticality safety in support of aqueous chloride operations. The sensitivity to the chlorine capture cross section is orders of magnitude greater than other existing benchmarks. The current limits, as defined by nuclear criticality safety, are 520 g Pu per batch, i.e. the minimum critical mass of the Pu solution infinitely reflected by water [Criticality Handbook: Volume II, (1969)]. This extremely conservative critical mass limit does not credit any neutron capture by chlorine (in particular neutron capture by 35 Cl) and greatly impedes the throughput required for current and future operations.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Development of an Improved RELAP5-3D Model for the High Temperature Test Facility

High-temperature gas-cooled reactors (HTGRs) are rapidly approaching deployment. Confidence in transient analysis of these systems for design, optimization, and licensing calculations requires modeling and simulation tools that have been validated against data relevant to HTGR conditions. The High Temperature Test Facility (HTTF) is an integral effects thermal hydraulics test facility for prismatic HTGRs. In spring and summer of 2019, HTTF was used for a series of experiments that now serve as the basis for the OECD/NEA Thermal Hydraulic Code Validation Benchmark for High Temperature Gas-Cooled Reactors using HTTF Data (HTGR T/H Benchmark). This benchmark contains problems for systems code, computational fluid dynamics (CFD), and coupled systems code/CFD modeling representing lower plenum mixing and both the depressurized and pressurized conduction cooldown (DCC and PCC respectively) transients. Benchmark problems include exercises for code-to-code and code-to-data comparisons as well as an exercise for error scaling between HTTF and the Modular High Temperature Gas-Cooled Reactor, which serves as the basis for the HTTF design. Previous analysis as part of the HTGR T/H benchmark used a RELAP5-3D model developed at Idaho National Laboratory (INL) and demonstrated an ability to reproduce trends in the measured data but difficulties reproducing experimental values within their uncertainty. These difficulties were largely attributed to assumptions made during the development of the initial RELAP5-3D model, which predated the HTTF experiments. A significant cause of difficulty reproducing the measured temperatures may be the radial nodalization of the previous RELAP5-3D model. The new model provides a finer nodalization to assess the impact of radial nodalization and allows for asymmetric heating within the core, which was a feature of multiple HTTF experiments. In this paper, we present the new RELAP5-3D model of HTTF. In addition to describing the new model, this paper compares the new and old models and provides results for a full-power steady state, a DCC, and a PCC in HTTF. These analyses are based on the code-to-code comparison exercises for the DCC and PCC problems of the HTGR T/H benchmark. We present the results of these exercises from the new model and compare them to the results of the old model.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Integrated Energy-Water Data for Cross-Sector Resilience

This white paper focuses on the “energy-for-water” domain, addressing the urgent need for integrated, empirical data to support regional management, benchmarking, and research on improving efficiency and developing technologies for water and wastewater management systems. The costs and energy required for the supply, treatment, and distribution of water and wastewater lack a standard data collection mechanism and centralized database or storage infrastructure, limiting data-driven decision-making across interdependent infrastructure systems.

42 ENGINEERING↗

JARVIS-Leaderboard: a large scale benchmark of materials design methods

Abstract Lack of rigorous reproducibility and validation are significant hurdles for scientific development across many fields. Materials science, in particular, encompasses a variety of experimental and theoretical approaches that require careful benchmarking. Leaderboard efforts have been developed previously to mitigate these issues. However, a comprehensive comparison and benchmarking on an integrated platform with multiple data modalities with perfect and defect materials data is still lacking. This work introduces JARVIS-Leaderboard, an open-source and community-driven platform that facilitates benchmarking and enhances reproducibility. The platform allows users to set up benchmarks with custom tasks and enables contributions in the form of dataset, code, and meta-data submissions. We cover the following materials design categories: Artificial Intelligence (AI), Electronic Structure (ES), Force-fields (FF), Quantum Computation (QC), and Experiments (EXP). For AI, we cover several types of input data, including atomic structures, atomistic images, spectra, and text. For ES, we consider multiple ES approaches, software packages, pseudopotentials, materials, and properties, comparing results to experiment. For FF, we compare multiple approaches for material property predictions. For QC, we benchmark Hamiltonian simulations using various quantum algorithms and circuits. Finally, for experiments, we use the inter-laboratory approach to establish benchmarks. There are 1281 contributions to 274 benchmarks using 152 methods with more than 8 million data points, and the leaderboard is continuously expanding. The JARVIS-Leaderboard is available at the website: https://pages.nist.gov/jarvis_leaderboard/

36 MATERIALS SCIENCE↗

Demography, dynamics and data: building confidence for simulating changes in the world's forests

Vegetation demographic models (VDMs) are advanced tools for simulating forest responses to climate and land-use changes, and are essential for projecting carbon cycling and large-scale forest management strategies. Despite their increasing incorporation into Earth System Models, VDMs differ in their demographic assumptions, with no prior quantitative comparison of their performance. We benchmarked nine VDMs against observational data from boreal, temperate and tropical sites, assessing their accuracy in predicting tree growth, carbon turnover, biomass stocks and size distributions. Models were simulated under consistent climate conditions with postdisturbance recovery monitored for at least 420 yr. Postdisturbance carbon recovery trajectories showed significant variability while remaining within observational ranges. Initial regrowth rates varied substantially (0.03-0.60, 0.18-0.70 and 0.35-1.10 kgCm-2 yr-1 for boreal, temperate and tropical sites, respectively), influenced by each model's initial forest state. Models captured mature forest carbon content but showed compensating effects between overestimated growth and underestimated mortality rates. This first multi-model benchmarking identifies growth and mortality rates as critical calibration targets and highlights the need to refine postdisturbance establishment conditions for model development. We outline specific benchmarking variables needed to improve predictions of forest responses to environmental change.

demographic vegetation model benchmarking↗

FY25 Mid-Year Report: FNCL Enhancements Implementation

During the first half of FY25 the FNCL team has made consistent progress toward the completion of our project goals. The FNCL prototype panel design has been successfully applied to a fully instrumented 3-panel system which is actively under construction. The FNCL Demonstrator System contains solid scintillators instrumented with SiPMs, which operate on an updated CAEN digitizer, requires no high-voltage, and has a smaller overall footprint. The onboard software will include the LLNL-developed GMM-PSD signal processing. Later this year the system will be experimentally tested alongside the baseline FNCL instrument at LLNLs ISSA facility. In addition to a full systems test, the performance of a DD generator for active interrogation measurements compared to the standard AmLi source will be established for both systems. The data collected at the ISSA facility will be used to experimentally validate the FNCL-Fast Isotopic Fuel Assay’s (FIFA) capability to measure U-235 loading and to predict gadolinium poison content with passive interrogation. The FNCL-FIFA modal was benchmarked with simulation-based data and a user-friendly GUI was added earlier this year. Three separate codes have been submitted to the LLNL ESW system for review prior to their transfers. These include the Predictive Modeling Response toolkit, GMM-PSD firmware beta version, and the FNCL-FIFA analysis package with GUI and user documentation.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Azimuthal correlation anisotropies in p + p collisions simulated using Pythia

Stimulated by a keen interest in possible collective behavior in high-energy proton-proton and proton-nucleus collisions, we study two-particle angular correlations in pseudorapidity and azimuthal differences in simulated p + p interactions using the Pythia 8 event generator. Multi-parton interactions and color connection are included in these simulations, which have been perceived to produce collectivity in final-state particles. Meanwhile, contributions from genuine few-body nonflow correlations, not of collective flow behavior, are known to be severe in these small-system collisions. We present our Pythia correlation studies pedagogically and report azimuthal harmonic anisotropies analyzed using several methods. We observe anisotropies in these Pythia simulated events qualitatively and semi-quantitatively, similar to experimental data. Furthermore, our findings highlight the delicate nature of azimuthal anisotropies in small-system collisions and provide a benchmark that can aid in improving data analysis and interpreting experimental measurements in small-system collisions.

Pythia↗

Constraining the impact of chlorine as a neutron absorber in next-gen fast reactor designs

The role of chlorine as a neutron poison and as a seed for producing radioactive waste in nuclear systems has driven a renewed interest to improve its nuclear data uncertainties. Additionally, basic and applied science programs that use CLYC (Cs 2 LiYCl 6 :Ce) detectors for neutron spectroscopy and monitoring are also very sensitive to any change in chlorine nuclear data for simulations of the detector response. In this work, sensitivities relevant for these different applications are addressed through simulations of the efficiency of CLYC detectors in a fast fission spectrum when applying new chlorine nuclear data as input. These simulations are validated by an experimental measurement using CLYC detectors coupled to an ionization chamber loaded with a 252 Cf spontaneous fission source. The results are then used to obtain the first reliable direct measurement of the 35 Cl(n,p 0 ) and summed Cl(n,p+n,α) fission spectrum average cross sections, found to be 54.7(32) and 105.0(98) mb, respectively. The results are within uncertainty of calculated fission spectrum averaged cross sections based on recently re-evaluated chlorine nuclear data, which confirm recent impact studies performed for the Molten Chloride Reactor Experiment. Meanwhile, there currently exists only one published criticality benchmark experiment that is sufficiently sensitive to chlorine nuclear data. Discrepancies are found with this set of criticality safety benchmarks, which are more sensitive to thermal and epithermal neutron energies than the energies, above 100 keV, tested in this current work. Hence, there is still a need to re-evaluate the chlorine nuclear data at lower energies to assess these discrepancies. Interpretation of the data from future “faster” criticality benchmarks, which are needed for next-gen fast reactor designs, benefit from the improved constraints on the chlorine nuclear data validated in this work.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Q1-2025 Solar Photovoltaic System Cost Benchmarks

The U.S. Department of Energy’s solar office and its national laboratory partners analyze cost data for U.S. solar photovoltaic systems to develop cost benchmarks to measure progress toward goals and guide research and development programs. The cost data is available in an Excel spreadsheet model that can be downloaded here.

14 SOLAR ENERGY↗

Benchmarking universal machine learning interatomic potentials for rapid analysis of inelastic neutron scattering data

The accurate calculation of phonons and vibrational spectra remains a significant challenge, requiring highly precise evaluations of interatomic forces. Traditional methods based on the quantum description of the electronic structure, while widely used, are computationally expensive and demand substantial expertise. Emerging universal machine learning interatomic potentials (uMLIPs) offer a transformative alternative by employing pre-trained neural network surrogates to predict interatomic forces directly from atomic coordinates. This approach dramatically reduces computation time and minimizes the need for technical knowledge. In this paper, we produce a phonon database comprising nearly 5000 inorganic crystals to benchmark the performance of several leading uMLIPs. We further assess these models in real-world applications by using them to analyze experimental inelastic neutron scattering data collected on a variety of materials. Through detailed comparisons, we identify the strengths and limitations of these uMLIPs, providing insights into their accuracy and suitability for fast calculations of phonons and related properties, as well as the potential for real-time interpretation of neutron scattering spectra. Our findings highlight how the rapid advancement of AI in science is revolutionizing experimental research and data analysis.

inelastic neutron scattering↗

Development of an Improved RELAP5-3D Model for the High Temperature Test Facility

High-temperature gas-cooled reactors (HTGRs) are rapidly approaching deployment. Confidence in transient analysis of these systems requires modeling and simulation tools that have been validated against data relevant to HTGR conditions. The High Temperature Test Facility (HTTF) is an integral effects thermal hydraulics test facility for prismatic HTGRs. In spring and summer of 2019, HTTF was used for a series of experiments that now serve as the basis for the Organization of Economic Cooperation and Development / Nuclear Energy Agency Thermal Hydraulic Code Validation Benchmark for High Temperature Gas-Cooled Reactors using HTTF Data (HTGR T/H Benchmark). Previous analyses as part of the HTGR T/H benchmark used a RELAP5-3D model developed at Idaho National Laboratory (INL) and demonstrated an ability to reproduce trends in the measured data but difficulties reproducing experimental values within their uncertainty. These difficulties were largely attributed to assumptions made during the development of the initial RELAP5-3D model, which predated the HTTF experiments. A significant cause of difficulty reproducing the measured temperatures may be the radial nodalization of the previous RELAP5-3D model. In this paper, we present a new RELAP5-3D model of HTTF with finer radial nodalization built to assess the impact of radial heat transfer. We describe the new model and compare it against the old one at full-power steady state and for the pressurized conduction cooldown (PCC) transient. These analyses are based on the code-to-code comparison exercise for the PCC problem of the HTGR T/H benchmark. We compare maximum block temperature as the primary figure of merit and include discussion on intracore natural circulation.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

OReole-FM: successes and challenges toward billion-parameter foundation models for high-resolution satellite imagery

While the pretraining of Foundation Models (FMs) for remote sensing (RS) imagery is on the rise, models remain restricted to a few hundred million parameters. Scaling models to billions of parameters has been shown to yield unprecedented benefits including emergent abilities, but requires data scaling and computing resources typically not available outside industry R&D labs. In this work, we pair high-performance computing resources including Frontier supercomputer, America's first exascale system, and high-resolution optical RS data to pretrain billion-scale FMs. Our study assesses performance of different pretrained variants of vision Transformers across image classification, semantic segmentation and object detection benchmarks, which highlight the importance of data scaling for effective model scaling. Moreover, we discuss construction of a novel TIU pretraining dataset, model initialization, with data and pretrained models intended for public release. By discussing technical challenges and details often lacking in the related literature, this work is intended to offer best practices to the geospatial community toward efficient training and benchmarking of larger FMs.

Ambrozio Dias, Philipe↗