Search NASA⌕ Search

SEARCH · Search NASA

Results for “Benchmark data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 559 records · Page 31

PRISMA: PARALLEL REFINEMENT AND INTEGRATION SYSTEM FOR MULTI-AZIMUTHAL ANALYSIS

The Parallel Refinement and Integration System for Multi-azimuthal Analysis (PRISMA, version 1.1.0) is a Python application for processing X-ray diffraction (XRD) image data. PRISMA wraps GSAS-II to perform azimuthally-binned peak refinement, computes per-frame strain and d-spacing from those fits, and provides three PyQt5 graphical interfaces: (1) a Recipe Builder for selecting GSAS-II control (.imctrl) files, optional mask (.immask) files or threshold-ased masking, reference and experiment image sets, peaks, zimuthal range and bin size, and an optional ceria-based auto-calibration; (2) a Batch Processor that uses Dask on local workstations and pure MPI (mpi4py.futures.MPICommExecutor) on HPC to distribute GSAS-II refinement across cores or compute nodes and write results to a 4-dimensional (peaks x frames x azimuths x measurements) Zarr dataset; and (3) a Data Analyzer that renders heatmaps of fit parameters, strain, frame-to-frame deltas, and percent-change-vs-reference, and exports user-defined subsections to CSV or Excel. The peak-refinement algorithm is deterministic. Benchmark on ALCF Crux: a 20,000-image set, single-peak fit in frame mode with 44 azimuthal bins on 128 nodes x 128 workers, 48 seconds total wall time.

Lorenzo Martin, Maria De La Cinta [Argonne Nationa↗

Identification of candidate host-specificity genes in Exserohilum turcicum using comparative genomics and transcriptomics

Abstract Exserohilum turcicum causes northern corn leaf blight and sorghum leaf blight. While the same species cause disease in both crops, the strains are host-specific. Here, we report the sequence and de novo annotated assemblies of one sorghum- and one maize-specific E. turcicum strain. The strains were sequenced using the PacBio Sequel II system. The total genome length for both assemblies was between 44 and 45 Mb with N50 of ∼2.5 Mb. Ninety-eight percent of the Benchmarking Universal Single-Copy Orthologs (BUSCO) for both assemblies had complete status. The estimated number of genes was 11,762 and 12,029 in the sorghum- and maize-specific isolates, respectively. Funannotate, EffectorP, SignalP, and transcriptome data were used to create functional annotation of each genome. The whole-genome comparison identified ten large-scale inversions and three translocations between the maize- and sorghum-specific strains, along with homologous genes and gene duplications. RNA was sequenced from the maize- and sorghum-specific isolate 10 days post-inoculation in maize and sorghum and from axenic cultures. Gene expression data from planta and axenic growth experiments were compared for each strain. Candidate host-specificity genes were identified by combining results from whole-genome comparison, synteny analysis, gene annotations, and transcriptome data. Overall, this study identified several candidate host-specificity genes that provide insights into E. turcicum interaction with its hosts.

Krone, Mara J. (ORCID:0000000159006624)↗

FATHOMS-RAG: A Framework for the Assessment of Thinking and Observation in Multimodal Systems that use Retrieval Augmented Generation

Retrieval-augmented generation (RAG) has emerged as a promising paradigm for improving factual accuracy in large language models (LLMs). We introduce a benchmark designed to evaluate RAG pipelines as a whole, evaluating a pipelines ability to ingest several modalities of information. We present (1) a curated dataset of 93 questions designed to evaluate a pipeline's ability to ingest textual data, tables, images, multimodal data, and cross-document multimodal data; (2) a phrase-level recall metric for correctness; (3) a nearest-neighbor embedding classifier in an attempt to classify pipeline hallucinations; (4) a comparative evaluation of 2 pipelines built with open-source retrieval mechanisms and 4 closed-source foundational models; and (5) a third-party human evaluation of the alignment of our correctness and hallucination metrics. We find that closed-source pipelines significantly outperform open-source pipelines in both the correctness and halucination metrics, with a wider performance gap in questions relying on multimodal and cross-document information. We also find after a human evaluation of our correctness and hallucination metric compared with our questions and pipeline responses, average agreement was 4.62 for correctness 4.53 for hallucination detection on a 1-5 Likert scale with 5 being strongly agree with our determination.

Hildebrand, Samuel [ORNL] (ORCID:0009000465963104)↗

Powering the Woods Hole X-Spar Buoy with Ocean Wave Energy—A Control Co-Design Feasibility Study

Despite its success in measuring air–sea exchange, the Woods Hole Oceanographic Institution’s (WHOI) X-Spar Buoy faces operational limitations due to energy constraints, motivating the integration of an energy harvesting apparatus to improve its deployment duration and capabilities. This work explores the feasibility of an augmented, self-powered system in two parts. Part 1 presents the collaborative design between X-Spar developers and wave energy researchers translating user needs into specific functional requirements. Based on requirements like desired power levels, deployability, survivability, and minimal interference with environmental data collection, unsuitable concepts are pre-eliminated from further feasibility study consideration. In part 2, we focus on one of the promising concepts: an internal rigid body wave energy converter. We apply control co-design methods to consider commercial of the shelf hardware components in the dynamic models and investigate the concept’s power conversion capabilities using linear 2-port wave-to-wire models with concurrently optimized control algorithms that are distinct for every considered hardware configuration. During this feasibility study we utilize two different control algorithms, the numerically optimal (but acausal) benchmark and the optimized damping feedback. We assess the sensitivity of average power to variations in drive-train friction, a parameter with high uncertainty, and analyze stroke limitations to ensure operational constraints are met. Our results indicate that a well-designed power take-off (PTO) system could significantly extend the WEC-Spar’s mission by providing additional electrical power without compromising data quality.

autonomous systems↗

Hybrid Quantum–Classical Graph Transformers for Efficient Sentiment Analysis

Quantum Machine Learning (QML) offers a promising paradigm that leverages quantum computing principles to develop efficient and expressive models for learning from complex and structured data. Recent advances in natural language processing (NLP) and artificial intelligence (AI) have demonstrated capabilities in understanding, generating, and reasoning over linguistic and multimodal information. In this work, we present the Quantum Graph Transformer (QGT), a hybrid quantum–classical architecture that extends graph transformer capabilities through quantum self-attention. The QGT models variable-length sentences as token graphs, where both the embedding encoding and the self-attention mechanisms are implemented using parameterized quantum circuits (PQCs), enabling efficient contextual learning with significantly fewer trainable parameters. We train QGT using both fully connected and 𝑘 -nearest-neighbor graph structures and evaluate it on five benchmark sentiment-classification datasets. Experimental results show that QGT consistently achieves higher or comparable accuracy to existing quantum NLP models and outperforms a Classical Graph Transformer (CGT) baseline with identical architecture, achieving 29.4 × fewer parameters while requiring 3–5 × fewer samples to reach comparable performance. These findings highlight the potential of graph-based quantum models as scalable and data-efficient architectures for natural language understanding.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Battery Electrolyte Design for Electric Vertical Takeoff and Landing (eVTOL) Platforms

Here, the development of robust and high-performance battery systems is crucial for the advancement of Electric Vertical Takeoff and Landing (eVTOL) vehicles for urban air mobility. This study evaluates the performance of different lithium-ion battery chemistries under Electric Vertical Takeoff and Landing (eVTOL) load profiles. The actual flight data coupled with physical models is used to create discharge profiles for testing on developed lithium-ion cells for eVTOLs. The performance of a standard liquid electrolyte (1.2 M LiPF 6 in EC:EMC), labeled Gen-2, is benchmarked and compared with a fast-charging electrolyte (1.2 M LiFSI in EC:EMC), labeled XFC. Cell analysis involves the use of various techniques, such as impedance spectroscopy, polarization curves, and capacity retention measurements. Capacity retention is stable for both systems over 500 cycles, but unique discharge capacity trends are observed for different mission segments. During the initial takeoff hover stages, Gen-2 electrolytes experience substantial voltage fade, while XFC electrolytes maintain consistent behavior. In general, the Gen-2 electrolyte demonstrated lower discharge overpotentials and higher decay during cycling compared to the XFC electrolyte. This work highlights the complexity of eVTOL battery behavior and provides insights into battery system design, contributing to the advancement of battery energy storage solutions for urban air mobility.

25 ENERGY STORAGE↗

Ab initio investigation of a hypersonic double cone experiment

This article presents a direct molecular simulation (DMS) of a reactive Mach 8.2 oxygen flow over a double cone geometry. The free stream conditions and article configuration generate a flow with thermal and chemical nonequilibrium, which are common attributes of hypersonic flight. This scenario was first studied experimentally at Calspan University of Buffalo Research Center’s test facility. DMS is a particle method that uses quantum mechanically derived interaction potentials to simulate molecular collisions within a flow field. Since these interaction potentials are the only modeling inputs used in the simulation, all flow features can solely be attributed to the ab initio potential energy surfaces. Hence, providing a comparison of a hypersonic ground test and numerical data anchored to quantum mechanics. The fundamental nature of DMS is leveraged to investigate molecular level mechanisms prevalent in the flow, and comparisons with lower fidelity simulations are presented to highlight the role of these first principles calculations as benchmark solutions.

Science & Technology - Other Topics↗

Variance Preserving Spectral Subsampling

Generating statistically faithful short-duration gamma-ray spectra from a single long measurement is essential in nuclear safeguards, supporting tasks such as algorithm development and machine-learning applications, especially when list-mode data are unavailable. Existing subsampling methods often distort the statistical characteristics of genuine short-duration measurements, leading to biased or unreliable analytical outcomes and thereby undermining downstream tasks. In this work, we compare five subsampling approaches using a benchmark set of 156 genuine replicate spectra collected with a high-purity germanium detector. We evaluate each method with respect to run-to-run variance, channel-to-channel variance, and preservation of total counts (losslessness). Across a wide range of subsampling ratios, only binomial subsampling without replacement consistently reproduces the statistical properties of genuine short-duration spectra, maintaining proper dispersion even in sparse spectral regions and perfectly preserving total counts. These results provide a mathematically principled and practically validated framework for generating synthetically shortened spectra when true short-duration measurements are unavailable.

98 NUCLEAR DISARMAMENT, SAFEGUARDS, AND PHYSICAL P↗

(Doublon) Benchmarking of Different Inverse Point Kinetics Implementations for an Autocorrected Reactimeter Algorithm

In November 2017, the Transient Reactor Test Facility returned to operation. Since that time, many transient test series have been completed, such as the Transient Heatsink Overpower Response capsule (THOR), the Transient Water Irradiation System for TREAT (TWIST), and Sirius. Each has provided valuable data for materials performance and reactor safety that can be applied in future designs. During each experimental series, detector count rates provided important information on the core behavior during transients. However, a limitation of these data is that variations in the neutron distribution during experiments can cause errors when attempting to infer reactivity evolution from detector signals. Neutron physics codes can be used to compute the flux shape variations. However, this is a poor solution when the experimental data is used for code verification, validation and uncertainty quantification. Indeed, if the output of the code is used both as a reference and to correct what the reference is compared to, the circular dependency limits the quality of the verification, validation and uncertainty quantification approach. To overcome this problem, the autocorrected reactimeter algorithm (ACRA) has been developed. This approach infers a time-dependent reactivity evolution by testing different spatial corrections and selecting the one that minimizes reactivity variations when the core is in a frozen configuration (i.e., when there is no variation in parameters affecting reactivity). However, the scope of this method was limited to transients where there were negligible thermal feedback. Indeed, the core is never in a frozen configuration when the fuel temperature varies during the whole transient. This is our motivation for developing an improved version of the ACRA that does not require frozen configurations. To develop this new algorithm, we need a precise and unbiased implementation of the inverse point kinetic equations (IPKEs) as any error in the reactivity evaluation will be propagated into the choice of the optimal spatial correction. Indeed, the previous reactimeter algorithm would use approximations, such as a negligible flux amplitude derivative, to focus on rapidity. For the numerical validation of ACRA, we aim at absolute error under for reactivity derived from signals similar to the one of this study. In this summary, we test eight different IPKE implementations. Each will process a mockup signal built for this study, similar to those that the future ACRA will process. Each reactivity output will be compared to the reference reactivity that has been used to generate the mockup signal. The implementation minimizing the difference with the reference reactivity will be used in the development of a new ACRA formulation.

73 - NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Search for single production of vector-like quarks decaying into W(ℓν)b in pp collisions at $\sqrt{s}=13$ TeV with the ATLAS detector

A search for single production of a vector-like quark Q, which could be either a singlet T, with charge $\frac{2}{3}$, or a Y from a (T, B, Y) triplet, with charge $-\frac{4}{3}$, is performed using data from proton-proton collisions at a centre-of-mass energy of 13 TeV. The data correspond to the full integrated luminosity of 140 fb −1 recorded with the ATLAS detector during Run 2 of the Large Hadron Collider. The analysis targets Q → Wb decays where the W boson decays leptonically. The data are found to be consistent with the expected Standard Model background, so upper limits are set on the cross-section times branching ratio, and on the coupling of the Q to the Standard Model sector for these two benchmark models. Effects of interference with the Standard Model background are taken into account. For the singlet T, the 95% confidence level limit on the coupling strength κ ranges between 0.22 and 0.52 for masses from 1150 to 2300 GeV. For the (T, B, Y) triplet, the limits on κ vary from 0.14 to 0.46 for masses from 1150 to 2600 GeV.

beyond Standard Model↗

Kernel methods for evolution of generalized parton distributions

Generalized parton distributions (GPDs) characterize the 3-dimensional structure of hadrons, combining information about their internal quark and gluon longitudinal momentum distributions and transverse position within the hadron. The dependence of GPDs on the factorization scale Q 2 allows one to connect hard exclusive processes involving GPDs at disparate energy and momentum scales, which is needed in global analyses of experimental data. Here, in this work, we explore how finite element methods can be used to construct fast and differentiable Q 2 evolution codes for GPDs in momentum space, which can be used in a machine learning framework. We show numerical benchmarks of the methods' accuracy, including a comparison to an existing evolution code from PARTONS/APFEL++, and provide a repository where the code can be accessed.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

The design space of E(3)-equivariant atom-centred interatomic potentials

Abstract Molecular dynamics simulation is an important tool in computational materials science and chemistry, and in the past decade it has been revolutionized by machine learning. This rapid progress in machine learning interatomic potentials has produced a number of new architectures in just the past few years. Particularly notable among these are the atomic cluster expansion, which unified many of the earlier ideas around atom-density-based descriptors, and Neural Equivariant Interatomic Potentials (NequIP), a message-passing neural network with equivariant features that exhibited state-of-the-art accuracy at the time. Here we construct a mathematical framework that unifies these models: atomic cluster expansion is extended and recast as one layer of a multi-layer architecture, while the linearized version of NequIP is understood as a particular sparsification of a much larger polynomial model. Our framework also provides a practical tool for systematically probing different choices in this unified design space. An ablation study of NequIP, via a set of experiments looking at in- and out-of-domain accuracy and smooth extrapolation very far from the training data, sheds some light on which design choices are critical to achieving high accuracy. A much-simplified version of NequIP, which we call BOTnet (for body-ordered tensor network), has an interpretable architecture and maintains its accuracy on benchmark datasets.

Computer Science↗

A combinatorially complete epistatic fitness landscape in an enzyme active site

Protein engineering often targets binding pockets or active sites which are enriched in epistasis—nonadditive interactions between amino acid substitutions—and where the combined effects of multiple single substitutions are difficult to predict. Few existing sequence-fitness datasets capture epistasis at large scale, especially for enzyme catalysis, limiting the development and assessment of model-guided enzyme engineering approaches. We present here a combinatorially complete, 160,000-variant fitness landscape across four residues in the active site of an enzyme. Assaying the native reaction of a thermostable β-subunit of tryptophan synthase (TrpB) in a nonnative environment yielded a landscape characterized by significant epistasis and many local optima. These effects prevent simulated directed evolution approaches from efficiently reaching the global optimum. There is nonetheless wide variability in the effectiveness of different directed evolution approaches, which together provide experimental benchmarks for computational and machine learning workflows. The most-fit TrpB variants contain a substitution that is nearly absent in natural TrpB sequences—a result that conservation-based predictions would not capture. Thus, although fitness prediction using evolutionary data can enrich in more-active variants, these approaches struggle to identify and differentiate among the most-active variants, even for this near-native function. Overall, this work presents a large-scale testing ground for model-guided enzyme engineering and suggests that efficient navigation of epistatic fitness landscapes can be improved by advances in both machine learning and physical modeling.

biocatalysis↗

Algal Biomass Production via Open Pond Algae Farm Cultivation: 2023 State of Technology and Future Research

The annual State of Technology (SOT) assessment is an essential activity for platform research conducted under the Bioenergy Technologies Office (BETO). It allows for the impact of research progress (both directly achieved in-house at the National Renewable Energy Laboratory [NREL] and furnished by partner organizations) to be quantified in terms of economic improvements in the overall biofuel production process for a particular biomass processing pathway, whether based on terrestrial or algal biomass feedstocks. As such, initial benchmarks can be established for currently demonstrated performance, and progress can be tracked toward out-year goals to ultimately demonstrate economically viable biofuel technologies. NREL's algae SOT benchmarking efforts historically focused both on front-end algal biomass production and separately on back-end conversion to fuels through NREL's "combined algae processing" (CAP) pathway. The production model is based on outdoor long-term cultivation data, enabled by comprehensive algal biomass production trials conducted under the Development of Integrated Screening, Cultivar Optimization, and Verification Research (DISCOVR) consortium efforts, driven by data furnished by Arizona State University (ASU) at the Arizona Center for Algae Technology and Innovation (AzCATI) testbed site. The CAP model is based on experimental efforts conducted primarily under NREL research and development projects. This report focuses on front-end algal biomass production, documenting the pertinent algal biomass cultivation parameters that were input to the NREL open pond algae farm model. Through partnerships under DISCOVR, collaborators at ASU furnished details on cultivation performance metrics including biomass productivity and harvest densities for recent growth trials done at the AzCATI site. The resulting biomass productivity was calculated at 16.7 g/m 2 /day (ash-free dry weight [AFDW], annual average) for seasonal cultivation of Picochlorum celeri TG2 and Monoraphidium minutum 26B-AM biomass strains at the ASU site. Picochlorum celeri achieved the best productivity from April to September, with Monoraphidium minutum 26B-AM being used between October and March. Tetraselmis striata LANL1001, usually part of the strain rotation in previous cultivation SOTs, was supplanted by Monoraphidium minutum 26B-AM in this year's outdoor cultivation trials. Finally, building from an industry case study presented in the 2022 SOT report, in the Appendix of this report we provide an update on further improved data furnished by an industry collaborator and resultant impacts on economics reflecting several seasonal scenarios. This case study provides a supplementary datapoint on work being performed elsewhere with a more dedicated focus on improved compositional quality, producing biomass enriched in lipids as may be more optimal for conversion upgrading to fuels and products.

09 BIOMASS FUELS↗

ILAMBv2.7 benchmarking results comparing E3SMv2.1 land-atmosphere coupled (BGCv2LNDATM) and stand alone land (ELM) simulations with CMIP6 emission driven historical simulations

This dataset contains land model benchmarking results for the Energy Exascale Earth System Model version 2.1 (E3SMv2.1), including outputs from both coupled biogeochemistry simulations and stand-alone land model simulations. These results are compared against several emission-driven historical simulations from the Coupled Model Intercomparison Project Phase 6 (CMIP6). Benchmarking was conducted using the International Land Model Benchmarking (ILAMB) package, version 2.7 (ILAMBv2.7). CMIP6 model outputs were sourced from the Earth System Grid Federation (ESGF), while the E3SMv2.1 results were derived from raw model outputs. These outputs underwent processing steps such as time serialization, conservative regridding, and data standardization to ensure comparability. For spatial interpolation, the Earth System Modeling Framework (ESMF) tool, ESMF_RegridWeightGen, was employed to generate regridding weights, enabling the transformation of E3SM’s native cubed-sphere grid to a regular latitude-longitude grid.

Feng, Sha [PNNL]↗

LHCspin: a Polarized Gas Target for LHC

The goal of the LHCspin project is to develop innovative solutions for measuring the 3D structure of nucleons in high-energy polarized fixed-target collisions at LHC, exploring new processes and exploiting new probes in a unique, previously unexplored, kinematic regime. A precise multi-dimensional description of the hadron structure has, in fact, the potential to deepen our understanding of the strong interactions and to provide a much more precise framework for measuring both Standard Model and Beyond Standard Model observables. This ambitious task poses its basis on the recent experience with the successful installation and operation of the SMOG2 unpolarized gas target in front of the LHCb spectrometer. Besides allowing for interesting physics studies ranging from astrophysics to heavy-ion physics, SMOG2 provides an ideal benchmark for studying beam-target dynamics at the LHC and demonstrates the feasibility of simultaneous operation with beam-beam collisions. With the installation of the proposed polarized target system, LHCb will become the first experiment to simultaneously collect data from unpolarized beam-beam collisions at $\sqrt{s}$=14 TeV and polarized and unpolarized beam-target collisions at $\sqrt{s_{NN}}\sim$100 GeV. LHCspin has the potential to open new frontiers in physics by exploiting the capabilities of the world's most powerful collider and one of the most advanced spectrometers. This document also highlights the need to perform an R&D campaign and the commissioning of the apparatus at the LHC Interaction Region 4 during the Run 4, before its final installation in LHCb. This opportunity could also allow to undertake preliminary physics measurements with unprecedented conditions.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Gaining Real-Time Water Leak Detection

Devens Reserve Forces Training Area is a United States Army Reserve (USAR) Installation that struggles with severe water leaks, often causing significant damage to the facility and requiring major renovation. Traditional water use is highly dependent on occupancy, so it can be difficult to benchmark a facility’s water use. It can be exceptionally difficult when occupancy is transient and/or varies. Pacific Northwest National Laboratory (PNNL) collaborated with Devens to implement real-time monitoring of their water consumption by utilizing the smart meter data from their existing 23 water meters. PNNL created a simple algorithm to calculate hourly water consumption and trigger an alert to be instantly emailed to Devens’ personnel when there appears to be a water leak in any building with a smart water meter. Here, this approach is expected to save hundreds of thousands of dollars in unnecessary water consumption costs and damages from leaks and was implemented with little-to-no costs or service disruptions. Next steps for this project include slow leak detection through nighttime monitoring and to extrapolate this water leak approach to the remainder 360 water meters on USAR’s Enterprise Building Control System so USAR sites across the country can be instantly notified of potential water leaks.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

InverseBench: Inverse design benchmark suite that contains inverse problems from science and engineering (InverseBench) v0.0.1

A software package that contains three inverse design blackbox problems to investigate the efficiency and accuracy of inverse design machine learning models. The software contains highly accurate forward machine learning models that can be used to assess the inverse predictions. The package also contains separate test data for each problem. The inverse design problems that are in the package are: airfoil inverse design, scalar boundary reconstruction and photonic surfaces inverse design.

Grbcic, Luka [Lawrence Berkeley National Laborator↗