Search NASASearch

SEARCH · Search NASA

Results for “Benchmark data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

ENDF/B-VIII.1: Updated Nuclear Reaction Data Library for Science and Applications

The ENDF/B-VIII.1 library is the newest recommended evaluated nuclear data file by the Cross Section Evaluation Working Group (CSEWG) for use in nuclear science and technology applications, and incorporates advances made in the six years since the release of ENDF/B-VIII.0. Among key advances made are that the 239 Pu file was reevaluated by a joint international effort and that updated 16,18 O, 19 F, 28–30 Si, 50–54 Cr, 55 Mn, 54,56,57 Fe, 63,65 Cu, 139 La, 233,235,238 U, and 240,241 Pu neutron nuclear data from the IAEA coordinated INDEN collaboration were adopted. Over 60 neutron dosimetry cross sections were adopted from the IAEA's IRDFF-II library. In addition, the new library includes significant changes for 3 He, 6 Li, 9 Be, 51 V, 88 Sr, 103 Rh, 140,142 Ce, Dy, 181 Ta, Pt, 206–208 Pb, and 234,236 U neutron data, and new nuclear data for the photonuclear, charged-particle and atomic sublibraries. Numerous thermal neutron scattering kernels were reevaluated or provided for the very first time. On the covariance side, work was undertaken to introduce better uncertainty quantification standards and testing for nuclear data covariances. The significant effort to reevaluate important nuclides has reduced bias in the simulations of many integral experiments with particular progress noted for fluorine, copper, and stainless steel containing benchmarks. Data issues hindered the successful deployment of the previous ENDF/B-VIII.0 for commercial nuclear power applications in high burnup situations. These issues were addressed by improving the 238 U and 239,240,241 Pu evaluated data in the resonance region. The new library performance as a function of burnup is similar to the reference ENDF/B-VII.1 library.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS

The boundary layer on compressor cascade blades

The purpose of NASA Research Grant NSG-3264 is to characterize the flowfield about an airfoil in a cascade at chord Reynolds number(R sub C)near 5 x 10 to the 5th power. The program is experimental and combines laser Doppler velocimeter (LDV) measurements with flow visualization techniques in order to obtain detailed flow data, e.g., boundary layer profiles, points of separation and the transition zone, on a cascade of highly-loaded compressor blades. The information provided by this study is to serve as benchmark data for the evaluation of current and future compressor cascade predictive models, in this way aiding in the compressor design process. Summarized is the research activity for the period 1 December 1985 through 1 June 1986. Progress made from 1 June 1979 through 1 December 1985 is presented. Detailed measurements have been completed at the initial cascade angle of 53 deg. (incidence angle 5 degrees). A three part study, based on that data, has been accepted as part of the 1986 Gas Turbine Conference and will be submitted for subsequent journal publication. Also presented are data for a second cascade angle of 45 deg (an incidence angle of 3 degrees).

Deutsch, S.

Open Rotor Tone Shielding Methods for System Noise Assessments Using Multiple Databases

Advanced aircraft designs such as the hybrid wing body, in conjunction with open rotor engines, may allow for significant improvements in the environmental impact of aviation. System noise assessments allow for the prediction of the aircraft noise of such designs while they are still in the conceptual phase. Due to significant requirements of computational methods, these predictions still rely on experimental data to account for the interaction of the open rotor tones with the hybrid wing body airframe. Recently, multiple aircraft system noise assessments have been conducted for hybrid wing body designs with open rotor engines. These assessments utilized measured benchmark data from a Propulsion Airframe Aeroacoustic interaction effects test. The measured data demonstrated airframe shielding of open rotor tonal and broadband noise with legacy F7/A7 open rotor blades. Two methods are proposed for improving the use of these data on general open rotor designs in a system noise assessment. The first, direct difference, is a simple octave band subtraction which does not account for tone distribution within the rotor acoustic signal. The second, tone matching, is a higher-fidelity process incorporating additional physical aspects of the problem, where isolated rotor tones are matched by their directivity to determine tone-by-tone shielding. A case study is conducted with the two methods to assess how well each reproduces the measured data and identify the merits of each. Both methods perform similarly for system level results and successfully approach the experimental data for the case study. The tone matching method provides additional tools for assessing the quality of the match to the data set. Additionally, a potential path to improve the tone matching method is provided.

Bahr, Christopher J.

Uncertainty Quantification in GADRAS Inverse Modeling

The Gamma Detector Response and Analysis Software (GADRAS) package includes an inverse modeling tool that is helpful in identifying characteristics of unknown radioactive materials. Traditionally, uncertainties in this analysis were derived solely from measurement data quality and the fit of synthetic spectra. This paper aims to rigorously quantify additional sources of uncertainty, focusing on uncertainties arising from measurements being analyzed, Detector Response Function (DRF) characterization, and DRF extrapolation. Applying these findings to the BeRPBall benchmark data set, we demonstrated the impact of these uncertainties on plutonium and polyethylene estimates. The results underscore the importance of incorporating diverse uncertainty sources to enhance the accuracy and reliability of GADRAS’s inverse modeling capabilities.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND

Airborne imaging spectroscopy surveys of Arctic and boreal Alaska and northwestern Canada 2017–2023

Since 2015, NASA’s Arctic Boreal Vulnerability Experiment (ABoVE) has investigated how climate change impacts the vulnerability and/or resilience of the permafrost-affected ecosystems of Alaska and northwestern Canada. ABoVE conducted extensive surveys with the Next Generation Airborne Visible/Infrared Imaging Spectrometer (AVIRIS-NG) during 2017, 2018, 2019, and 2022 and with AVIRIS-3 in 2023 to characterize tundra, taiga, peatlands, and wetlands in unprecedented detail. The ABoVE AVIRIS dataset comprises ~1700 individual flight lines covering ~120,000 km 2 with nominal 5 m × 5 m spatial resolution. Data include individual transects to capture important gradients like the tundra-taiga ecotone and maps of up to 10,000 km 2 for key study areas like the Mackenzie Delta. The ABoVE AVIRIS surveys enable diverse ecosystem science, provide crucial benchmark data for validating retrievals from the PACE, PRISMA, and EnMAP satellite sensors and help prepare for the SBG and CHIME missions. This paper guides interested researchers to fully explore the ABoVE AVIRIS spectral imagery and complements our guide to the ABoVE airborne synthetic aperture radar surveys.

Miller, Charles E. [California Institute of Techno

Empirically-calibrated H100 node power models for accurate AI training energy estimation

Accurately quantifying the energy use of artificial intelligence (AI) training is critical for infrastructure planning, carbon accounting, and sustainable data center operation, but few studies have directly measured the power consumption of production workloads on contemporary hardware. By combining empirical measurements from Brookhaven National Laboratory during AI training on 8-graphics-processing-unit H100 systems with open-source benchmarking data, we develop statistical models relating computational intensity to node-level power consumption. We measure the gap between manufacturer-rated thermal design power (TDP) and actual power demand during AI training. Our analysis reveals that even computationally intensive workloads operate at only 76% of the 10.2 kW TDP rating. Our architecture-specific model, calibrated to floating-point operations, predicts energy consumption with 11.4% mean absolute percentage error, significantly outperforming TDP-based approaches (27%–37% error). We identified distinct power signatures between transformer and convolutional neural network architectures, with transformers showing characteristic fluctuations that may impact grid stability. These results provide a measurement-grounded basis for improving AI training energy estimates, enabling more reliable infrastructure sizing, cost projections, and environmental impact assessments.

Newkirk, Alex C

Multi‐Model Ensembles in Ecosystem Modeling: Challenges and Best Practices for Decision‐Making

Ecosystem models are increasingly central to the decision-making for environmental policy, conservation planning, and climate-related investments. Yet, the growing reliance on Multi-Model Ensembles (MMEs) of ecosystem models by practitioners and policymakers, sometimes under tight timelines and imperfect information, has frequently outpaced the scientific rigor required to ensure ensemble reliability. Here, MMEs refer to approaches that combine targeted predictions from multiple models with the expectation of improving robustness and quantifying predictive uncertainty. Poorly designed MMEs may create a false sense of confidence and lead to suboptimal policy and market decisions. This perspective argues that robust decision-making-relevant MMEs must be grounded on two pillars: (1) rigorous Model Intercomparison Projects (MIPs), which identify inter-model agreement and disagreement, characterize model uncertainties, and evaluate robustness with observationally based benchmarks—MIPs' diagnostic evaluation is so critical that it must be needed to drive MME's decision in model selection and weighting, especially when only a limited number of models available; and (2) co-design by both stakeholders and scientists to ensure that scenarios, metrics and uncertainty requirements provide decision-relevant information. Building upon the past success and lessons from the existing MIPs-MMEs efforts (e.g., climate/Earth system/crop), we derived the theoretical basis for MMEs, addressed their specific challenges in ecosystem modeling, and highlighted proper consideration of model numbers and diversity, risk of model inter-dependence, effective calibration of model parameters, possible overdue of some ecosystem model development, critical roles of open benchmark data across a wide range of conditions, and suggested use of Artificial Intelligence to support MIPs-MMEs. We highlighted the under-recognized opportunity for MIPs and MMEs to drive scientific progress and innovation through identifying better performing models, systematic benchmarking, feedback loops, and targeted model improvement. By following actionable best practice guidelines, MMEs can evolve from ad hoc aggregation of models into a trusted backbone of environmental policy and decision-making.

ecosystem modeling

Kernel PLS-SVC for Linear and Nonlinear Discrimination

A new methodology for discrimination is proposed. This is based on kernel orthonormalized partial least squares (PLS) dimensionality reduction of the original data space followed by support vector machines for classification. Close connection of orthonormalized PLS and Fisher's approach to linear discrimination or equivalently with canonical correlation analysis is described. This gives preference to use orthonormalized PLS over principal component analysis. Good behavior of the proposed method is demonstrated on 13 different benchmark data sets and on the real world problem of the classification finger movement periods versus non-movement periods based on electroencephalogram.

Rosipal, Roman

Flame-Vortex Interactions in Microgravity to Improve Models of Turbulent Combustion

A unique flame-vortex interaction experiment is being operated in microgravity in order to obtain fundamental data to assess the Theory of Flame Stretch which will be used to improve models of turbulent combustion. The experiment provides visual images of the physical process by which an individual eddy in a turbulent flow increases the flame surface area, changes the local flame propagation speed, and can extinguish the reaction. The high quality microgravity images provide benchmark data that are free from buoyancy effects. Results are used to assess Direct Numerical Simulations of Dr. K. Kailasanath at NRL, which were run for the same conditions.

Driscoll, James F.

An experimental study of characteristic combustion-driven flows for CFD validation

The application of laser-based diagnostic techniques has become commonplace to a wide variety of combustion problems. New insights into combustion phenomena at a level previously unattainable has been made possible by non-intrusive measurements of velocity, temperature, and species. However, due to the adverse conditions which exist inside rocket engines, relatively few studies have addressed these combustion environments. The high pressure, high speed, combusting environment in a rocket engine prohibits the application of several measurement techniques. However, in the rocket community, there is a critical need for rocket flow field data to validate computational fluid dynamic (CFD) codes. Currently at Penn State, there is an effort to obtain flowfield measurements inside a rocket engine. Velocity measurements have been made inside the combustion chamber of a uni-element (shear coaxial injector) optically accessible rocket chamber at several axial locations downstream of the injector. These measurements, combined with future measurements, will provide benchmark data for CFD code validation.

Pal, S.

ENDF/B-VIII.1

The ENDF/B-VIII.1 release is the newest evaluated nuclear data library produced, distributed, and recommended by CSEWG for use in nuclear science and technology applications. Among the many key advances, relative to the previous version ENDF/B-VIII.0, are: re-evaluation of 239Pu file by a joint international effort; updated 16,18O, 19F, 28-30Si, 50-54Cr, 55Mn, 54,56,57Fe, 63,65Cu, 139La, 233,235,238U, and 240,241Pu neutron nuclear data by the IAEA-coordinated INDEN collaboration; significant changes for 3He, 6Li, 9Be, 51V, 88Sr, 103Rh, 140,142Ce, Dy, 181Ta, Pt, 206-208Pb, and 234,236U neutron data; new nuclear data for the photo-nuclear, being 196 adopted from the IAEA2019 Photonuclear Data Library and one new file from JENDL-5; and new evaluations for the charged-particle and atomic sublibraries. Numerous thermal neutron scattering kernels were re-evaluated or provided for the very first time. Additionally, new covariance testing was implemented. ENDF/B-VIII.1 reduced bias in the simulations of many integral experiments with particular progress noted for fluorine, copper and stainless steel containing benchmarks. Data issues which had hindered the deployment of ENDF/B-VIII.0 for commercial nuclear power applications in high burn-up situations, were addressed. ENDF/B-VIII.1 data are distributed in both ENDF-6 and GNDS formats.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS

Perspectives on Systematic Cloud Microphysics Scheme Development With Machine Learning

Cloud microphysics—the collection of processes that govern the small‐scale formation, evolution, and interactions of liquid droplets and ice crystals in clouds and precipitation—remains a major source of uncertainty in weather and climate models. Although too small in scale to be explicitly resolved in any large‐eddy simulation, weather, or climate model, the representation of cloud microphysical processes has significant impact at the climate scale. Current microphysical schemes are limited by both parametric uncertainty, linked to uncertainty in physical parameter values, and structural uncertainty, arising from incomplete physical understanding of the processes at play or approximations made for computational efficiency. Recent advances in the application of machine learning (ML) to the physical sciences show significant potential for minimizing these limitations by leveraging high‐fidelity simulations and observations. Here we outline the challenges that must be addressed to apply ML toward cloud microphysics scheme development. This perspectives paper synthesizes recent progress in using data‐driven methods, including ML, to improve cloud microphysics parameterizations and highlights opportunities to address key uncertainties. We discuss the roles of aleatoric (irreducible, or statistical) and epistemic (reducible, or systematic) errors in contributing to microphysics parameterization uncertainty. ML can leverage observations to improve microphysical schemes via bottom‐up and top‐down constraints. Methods such as differentiable programming and ML‐enhanced sampling strategies and the creation of large scale benchmark data sets promise to bridge the gap between observations and models and to improve the consistency of cloud microphysical representation across temporal and spatial scales.

Lamb, Kara D. [Columbia Univ., New York, NY (Unite

A format for the interchange of scheduling models

In recent years a variety of space-activity schedulers have been developed within the aerospace community. Space-activity schedulers are characterized by their need to handle large numbers of activities which are time-window constrained and make high demands on many scarce resources, but are minimally constrained by predecessor/successor requirements or critical paths. Two needs to exchange data between these schedulers have materialized. First, there is significant interest in comparing and evaluating the different scheduling engines to ensure that the best technology is applied to each scheduling endeavor. Second, there is a developing requirement to divide a single scheduling task among different sites, each using a different scheduler. In fact, the scheduling task for International Space Station Alpha (ISSA) will be distributed among NASA centers and among the international partners. The format used to interchange scheduling data for ISSA will likely use a growth version of the format discussed in this paper. The model interchange format (or MIF, pronounced as one syllable) discussed in this paper is a robust solution to the need to interchange scheduling requirements for space activities. It is highly extensible, human-readable, and can be generated or edited with common text editors. It also serves well the need to support a 'benchmark' data case which can be delivered on any computer platform.

Jaap, John P.

Extending quantum-mechanical benchmark accuracy to biological ligand-pocket interactions

Predicting the binding affinity of ligands to protein pockets is key in the drug design pipeline. The flexibility of ligand-pocket motifs arises from a range of attractive and repulsive electronic interactions during binding. Accurately accounting for all interactions requires robust quantum-mechanical (QM) benchmarks, which are scarce for ligand-pocket systems. Additionally, disagreement between “gold standard” Coupled Cluster (CC) and Quantum Monte Carlo (QMC) methods casts doubt on many benchmarks for larger non-covalent systems. We introduce the “QUantum Interacting Dimer” (QUID) benchmark framework containing 170 non-covalent (non-)equilibrium systems modeling chemically and structurally diverse ligand-pocket motifs. Symmetry-adapted perturbation theory shows that QUID broadly covers non-covalent binding motifs and energetic contributions. Robust binding energies are obtained using complementary CC and QMC methods, achieving agreement of 0.5 kcal/mol. The benchmark data analysis reveals that several dispersion-inclusive density functional approximations provide accurate energy predictions, though their atomic van der Waals forces differ in magnitude and orientation. Contrarily, semiempirical methods and empirical force fields require improvements in capturing non-covalent interactions (NCIs) for out-of-equilibrium geometries. The wide span of NCIs, highly accurate interaction energies, and analysis of molecular properties take QUID beyond the “gold standard” for QM benchmarks of ligand-protein systems.

Puleva, Mirela [University of Luxembourg, Luxembou

Characterizing and improving the performance of molten-salt-steam heat exchangers in concentrating solar power plants

Shell-and-tube heat exchangers (HXs) for steam generation from molten salts in concentrating solar power (CSP) plants experience thermal fatigue due to significant temperature gradients and inherent transient operation. Molten salt-steam HX design lifespans exceed actual lifespans, and, as a consequence, designers overpredict plant profitability and operators neglect appropriate prescriptions to optimize these lifetimes. Here, this study refines HX lifespan estimates with data benchmarked against thermal-fluid mechanical modeling of stress and accumulated fatigue. Reduced-order thermal models of the molten salt-steam, shell-and-tube evaporator and superheater predict transient temperature profiles along the two HXs salt-steam flow paths. The modeled evaporator and superheater temperature profiles enable assessment of cyclic stresses within the HX tubesheets, where molten-salt HX failures are most common. Evaporator and superheater performance data from a current 110 MW elec commercial CSP plant provide a basis for validating the reduced-order HX models. HX life predictions derived from stochastic failure distributions serve as inputs for simulating and optimizing existing plant operations. The impact of the updated lifespans on overall plant revenue depends on operating scenarios. This study suggests that typical ramping rates for a CSP plant with a high-temperature Rankine cycle result in an evaporator and superheater life of approximately 10 and 25 years, respectively, compared to the design target of 30 years. Reduced HX lifespans decrease operational plant revenue on average by 4.6-5.1%. Furthermore, there may be as many as four HX replacements over the 30-year lifetime of the plant; and, purchase agreement loss due to failure to meet contractual production requirements can have ramifications that include the risk of bankruptcy.

14 SOLAR ENERGY

Propanal, an Interstellar Aldehyde - First Infrared Band Strengths and Other Properties of the Amorphous and Crystalline Forms

Chemical evolution in molecular clouds in the interstellar medium is well established, with the identification of over 200 molecules and molecular ions. Among the classes of interstellar organic compounds found are the aldehydes. However, laboratory work on the aldehydes has scarcely kept pace with astronomical discoveries as little quantitative solid-phase infrared (IR) data has been published on any of the aldehydes, and the same is true for important properties such as density, refractive indices, and vapor pressures. In this paper we examine the IR spectra of solid propanal (HC(O)CH2CH3, propionaldehyde), along with several physical properties, for both the amorphous and crystalline forms of the compound. The quantitative measurements we report, such as infrared intensities and optical constants, will be useful in laboratory investigations of the formation and evolution of propanal-containing ices, will serve as benchmark data for theoretical investigations, and will inform observational studies.

Yukiko Yagi Yarnall

Building Lunar Maps for Terrain Relative Navigation and Hazard Detection Applications

Terrain Relative Navigation (TRN) systems localize a spacecraft with respect to a map of the surface by comparing descent imagery to that reference map. The spacecraft position estimates can only be as accurate as the reference map itself. Accurate map products that are based on orbital reconnaissance data must be validated for navigation applications to ensure that all relevant error sources are minimized. Currently available map products have been generated for scientific applications, so the need for accurate TRN maps remains a gap to be filled for upcoming lunar lander missions, in particular missions to the South Pole region. Additionally, representative high-resolution maps that contain lander-scale features are needed for successful development and testing of Hazard Detection (HD) systems. This paper describes one of NASA’s current efforts to develop benchmark data sets that can be used for developing and testing TRN and HD algorithms as well as suggested processes and metrics for generating and validating lunar maps that can be used for navigation and hazard detection.

Lunar Maps

Building Lunar Maps for Terrain Relative Navigation and Hazard Detection Applications

Terrain Relative Navigation (TRN) systems that localize a spacecraft with respect to a map of the surface by comparing descent imagery to that reference map can only be as accurate as the reference map itself. Accurate map products that are based on orbital reconnaissance data must be validated for navigation applications to ensure that all relevant error sources are minimized. Currently available map products have been generated for scientific applications, so the need for accurate TRN maps remains a gap to be filled for upcoming lunar lander missions, in particular missions to the South Pole region. Additionally, representative high-resolution maps that contain lander-scale features are needed for successful development and testing of Hazard Detection (HD) systems. This paper describes one of NASA’s current efforts to develop benchmark data sets that can be used for developing and testing TRN and HD algorithms as well as suggested processes and metrics for generating and validating lunar maps that can be used for navigation and hazard detection.

Beyer, Ross A.