Search NASASearch

SEARCH · Search NASA

Results for “Data usability”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

Bootstrap-determined p values in lattice QCD

We present a general method to determine the probability that stochastic Monte Carlo data, in particular those generated in a lattice QCD calculation, would have been obtained were that data drawn from the distribution predicted by a given theoretical hypothesis. Such a probability, or p -value, is often used as an important heuristic measure of the validity of that hypothesis. The proposed method offers the benefit that it remains usable in cases where the standard Hotelling T 2 methods based on the conventional χ 2 statistic do not apply, such as for uncorrelated fits. Specifically, we analyze q 2 , defined as the correlated χ 2 statistic obtained using an arbitrary covariance matrix estimator, and show how to use the bootstrap as a data-driven method to determine the expected distribution of q 2 for a given hypothesis with minimal assumptions. This distribution can then be used to determine the p -value for a fit to the data. We also describe a bootstrap approach for quantifying the impact upon this p -value of estimating population parameters from a single ensemble of N samples. The overall method is accurate up to a 1 / N bias which we do not attempt to quantify. Published by the American Physical Society 2025

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC

MolViewSpec: a Mol* extension for describing and sharing molecular visualizations

Data visualization is a pivotal component of a structural biologist’s arsenal. The Mol* Viewer makes molecular visualizations available to broader audiences via most web browsers. While Mol* provides a wide range of functionality, it has a steep learning curve and is only available via a JavaScript interface. To enhance the accessibility and usability of web-based molecular visualization, we introduce MolViewSpec (molstar.org/mol-view-spec), a standardized approach for defining molecular visualizations that decouples the definition of complex molecular scenes from their rendering. Scene definition can include references to commonly used structural, volumetric, and annotation data formats together with a description of how the data should be visualized and paired with optional annotations specifying colors, labels, measurements, and custom 3D geometries. Developed as an open standard, this solution paves the way for broader interoperability and support across different programming languages and molecular viewers, enabling more streamlined, standardized, and reproducible visual molecular analyses. MolViewSpec is freely available as a Mol* extension and a standalone Python package.

Midlik, Adam [European Bioinformatics Institute (U

Geothermal Power Systems Analysis: Outcome of Industry Stakeholders Workshop: Preprint

Geothermal cost and performance evaluation implemented via technoeconomic assessment (TEA) modeling is critical for the Department of Energy (DOE) and other geothermal industry stakeholders in assessing the current state of geothermal technologies and to identify existing hurdles to commercially viable geothermal development. The Geothermal Electricity Technology Evaluation Model (GETEM) is a major TEA tool used in estimating the economic feasibility and levelized cost of energy (LCOE) of conventional hydrothermal systems and enhanced geothermal systems (EGS). Since 2021, GETEM has been transitioning from an intricate spreadsheet model to a user-friendly tool within the System Advisor Model (SAM) developed by the National Renewable Energy Laboratory (NREL). Apart from enabling an expanded visibility of the geothermal model among other renewable resources, having GETEM in SAM has the advantage of simulation automation, better usability, updates tracking, active user inputs/feedback, and extended financial modeling. GETEM is used in developing supply curves for the Annual Technology Baseline (ATB). The ATB data are inputs to the Renewable Energy Potential (reV) and the Regional Energy Deployment System (ReEDS) models. The geothermal module in NREL’s reV model assesses the geothermal energy potential in the conterminous United States by defining the geospatial intersection of geothermal resources with existing grid infrastructure within the constraint of land use characteristics. The ReEDS model is a capacity expansion model used for simulating the long-term build-out and operation of the US generation and transmission system based on current energy costs and policies. To ensure enhanced representation of current industry trends in our model transitions and development, we organized a two-day virtual workshop to elicit geothermal industry stakeholder input and recommendations on our current approaches and assumptions on technoeconomic, resource assessment, and deployment scenarios modeling of geothermal technologies. Participants included developers, operators, investors, regulatory agencies, system modelers, national laboratory researchers, consultants, and other stakeholders. In this workshop, we gained stakeholder insights on current geothermal plant performance (i.e., capacity factors), updated drilling costs and learning curves, and next generation technologies such as closed loop and superhot rock geothermal. Other outcomes from this workshop and its impact on future geothermal development feasibility, resource availability, and capacity expansion studies are compiled and discussed.

Annual Technology Baseline

Laboratory-Based Micro-X-ray Computed Tomography of Energy Materials at Idaho National Laboratory

Abstract The Idaho National Laboratory (INL) has implemented laboratory-based micro-X-ray computed tomography in a laboratory equipped for the examination of highly radioactive samples. This capability provides nondestructive three-dimensional volumetric information on samples to inform subsequent traditional destructive examinations as well as real-world inputs for high-fidelity scientific modeling. Samples can be imaged with spatial resolutions ranging from several hundred nm/voxel up to ~ 100 µm/voxel. The best usable spatial resolution achieved to date is 384 nm/voxel with this instrument, while the highest radiological dose rate of a sample imaged is ~ 60 R/h β/γ on contact. Advanced data analysis, including custom tomographic reconstruction and segmentation methods, have also been developed to support this capability. In addition to traditional digital X-ray radiography and tomography, this instrument is also able to visualize in situ tensile and compression testing as well as perform diffraction contrast tomography. This work describes the X-ray computed tomography post-irradiation examination capabilities at INL, as well as detailing a variety of applications this instrument has examined.

36 MATERIALS SCIENCE

Shifts in evolutionary lability underlie independent gains and losses of root-nodule symbiosis in a single clade of plants

Abstract Root nodule symbiosis (RNS) is a complex trait that enables plants to access atmospheric nitrogen converted into usable forms through a mutualistic relationship with soil bacteria. Pinpointing the evolutionary origins of RNS is critical for understanding its genetic basis, but building this evolutionary context is complicated by data limitations and the intermittent presence of RNS in a single clade of ca. 30,000 species of flowering plants, i.e., the nitrogen-fixing clade (NFC). We developed the most extensive de novo phylogeny for the NFC and an RNS trait database to reconstruct the evolution of RNS. Our analysis identifies evolutionary rate heterogeneity associated with a two-step process: An ancestral precursor state transitioned to a more labile state from which RNS was rapidly gained at multiple points in the NFC. We illustrate how a two-step process could explain multiple independent gains and losses of RNS, contrary to recent hypotheses suggesting one gain and numerous losses, and suggest a broader phylogenetic and genetic scope may be required for genome-phenome mapping.

59 BASIC BIOLOGICAL SCIENCES

An Ab Initio Molecular Dynamics Study of Key Thermodynamic Input Parameters for Computer Simulation of U-6Nb Solidification

The key to metallic fuel development is the fabrication of uranium metal and alloys into fuel forms. U-Nb alloys are one of the best candidates for a metallic fuel alloy with high-temperature strength sufficient to support the core, acceptable nuclear properties, good fabricability, and compatibility with usable coolant media. Melt processing has been a key component of the metallic fuel cycle, and process models require thermophysical parameters at elevated temperatures, particularly above the melting temperatures, regarding which experimental data are scarce, for accurate simulations and process development. By means of ab initio density-functional theory (DFT) quantum molecular dynamics (QMD), we have calculated the main thermophysical parameters—the density, thermal expansion coefficient, specific heat, thermal conductivity, melting temperature, latent heat of fusion, and viscosity—used in the modeling of the U-6 wt.% Nb alloy casting. The melting temperature of the U-6 wt.% Nb alloy at ambient pressure is obtained by means of QMD simulations using the Z-method. The ambient volume change and latent heat of melting of U-6 wt.% Nb are also derived from QMD simulations in conjunction with analytical fitting for the energy and pressure. The thermal conductivity for the solid U-Nb alloy is calculated from the semi-classical Boltzmann transport equation combined with an estimate of the electron relaxation time obtained from DFT simulations.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS

Energy dataset of Frontier supercomputer for waste heat recovery

The Hewlett Packard Enterprise–Cray EX Frontier is the world’s first and fastest exascale supercomputer, hosted at the Oak Ridge Leadership Computing Facility in Tennessee, United States. Frontier is a significant electricity consumer, drawing 8–30 MW; this massive energy demand produces significant waste heat, requiring extensive cooling measures. Although harnessing this waste heat for campus heating is a sustainability goal at Oak Ridge National Laboratory (ORNL), the 30 °C–38 °C waste heat temperature poses compatibility issues with standard HVAC systems. Heat pump systems, prevalent in residential settings and some industries, can efficiently upgrade low-quality heat to usable energy for buildings. Thus, heat pump technology powered by renewable electricity offers an efficient, cost-effective solution for substantial waste heat recovery. However, a major challenge is the absence of benchmark data on high-performance computing (HPC) heat generation and waste heat profiles. This paper reports power demand and waste heat measurements from an ORNL HPC data centre, aiming to guide future research on optimizing waste heat recovery in large-scale data centres, especially those of HPC calibre.

97 MATHEMATICS AND COMPUTING

Machine learning materials properties with accurate predictions, uncertainty estimates, domain guidance, and persistent online accessibility

One compelling vision of the future of materials discovery and design involves the use of machine learning (ML) models to predict materials properties and then rapidly find materials tailored for specific applications. However, realizing this vision requires both providing detailed uncertainty quantification (model prediction errors and domain of applicability) and making models readily usable. At present, it is common practice in the community to assess ML model performance only in terms of prediction accuracy (e.g. mean absolute error), while neglecting detailed uncertainty quantification and robust model accessibility and usability. Here, we demonstrate a practical method for realizing both uncertainty and accessibility features with a large set of models. We develop random forest ML models for 33 materials properties spanning an array of data sources (computational and experimental) and property types (electrical, mechanical, thermodynamic, etc). All models have calibrated ensemble error bars to quantify prediction uncertainty and domain of applicability guidance enabled by kernel-density-estimate-based feature distance measures. All data and models are publicly hosted on the Garden-AI infrastructure, which provides an easy-to-use, persistent interface for model dissemination that permits models to be invoked with only a few lines of Python code. We demonstrate the power of this approach by using our models to conduct a fully ML-based materials discovery exercise to search for new stable, highly active perovskite oxide catalyst materials.

domain of applicability

Distributed Neural Representation for Reactive In Situ Visualization

Implicit neural representations (INRs) have emerged as a powerful tool for compressing large-scale volume data. This opens up new possibilities for in situ visualization. However, the efficient application of INRs to distributed data remains an underexplored area. Here, in this work, we develop a distributed volumetric neural representation and optimize it for in situ visualization. Our technique eliminates data exchanges between processes, achieving state-of-the-art compression speed, quality and ratios. Our technique also enables the implementation of an efficient strategy for caching large-scale simulation data in high temporal frequencies, further facilitating the use of reactive in situ visualization in a wider range of scientific problems. We integrate this system with the Ascent infrastructure and evaluate its performance and usability using real-world simulations.

Wu, Qi

A Reverse Logistics Tool For Ev Battery Recycling And Repurposing,

The demand for electric vehicles (EVs) in the United States is projected to rise significantly, with sales expected to reach approximately 4.1 million units by 2030. However, the U.S. remains heavily reliant on imports for the batteries and critical raw materials—such as lithium, cobalt, and nickel—that power these vehicles. As of 2024, around 70% of these imports originate from China. This dependency has become even more precarious following China’s imposition of export restrictions in April 2025, a retaliatory move against U.S. tariffs. These developments highlight the strategic vulnerabilities posed by China’s dominant position in the critical materials market. Compounding the issue, decades of intensive extraction have severely depleted global reserves of critical materials, widening the gap between supply and growing demand. This situation underscores the urgent need for the U.S. and other nations to diversify their sources of critical materials and enhance domestic capabilities to secure these resources—an essential step toward ensuring long-term energy security. At the end of their lifecycle—whether due to the battery’s degradation or the retirement of the vehicle—EV batteries are often improperly disposed of or sent to landfills. However, many of these batteries still retain usable capacity and can follow one of three alternative pathways: (a) Re-used: deployed in another vehicle with a shorter driving range, (b) Re-purposed: utilized act as a backup storage/power for data centers, solar panels, and e-scotters or (c) Recycled: broken down to recover the critical materials. To that end, the proposed tool (REBORN) is designed to optimize the reverse logistics network for battery repurposing and recycling. Its goal is to minimize associated costs while identifying optimal locations for battery collection and processing. Ultimately, REBORN ensures that each battery is used to its fullest potential.

Srinivas, SrikarV. [Idaho National Laboratory (INL

Evaluation of a Reduced-Order Model for IBR Fault Response Representation via OEM Blackbox Models

This paper presents a fully implemented inverter reduced-order-model (ROM) in an EMT simulation (PSCAD) library component for direct user utilization in protection studies. The developed inverter ROM has the following features: Equivalent to a full inverter-based resource (IBR) inverter model with positive- and negative-sequence current formulation and representation. A Python script is developed to fully automate this process, including training data generation, ROM parameter training, updating parameters, and model verification and validation. The ROM is validated using both IEEE 2800-compliant and non-compliant OEM modes in a real-world system, building confidence of its usability by protection engineers.

24 POWER TRANSMISSION AND DISTRIBUTION

Phoenix Dust Storm (PHX-DUST) Scale: A Cooperatively Developed Dust Storm Scale for Phoenix, Arizona

Using extensive localized meteorological and air quality data for central Arizona from 2010 to 2023, we create a postevent dust storm scale for use by the primary stakeholders in central Arizona for the Phoenix metropolitan area called the Phoenix Dust Storm (PHX-DUST) scale. To ensure usability across a wide spectrum of users, a core concept was to “keep it simple.” The PHX-DUST scale is based on (i) maximum dust concentration for particulate matter (PM10) across the network in micrograms per cubic meter which determines category, e.g., “category 5 (the highest observed category)” and category 4; (ii) the number of network monitors achieving dust concentrations above a threshold of 500 μg m−3 (categorized as “widespread” or “isolated,” as measures of spatial extent); (iii) a duration parameter (which determines “long duration” or “short duration”); and (iv) a measure of wind speed over the affected area (maximum wind gust recorded over the network, which may not be the maximum wind of the storm). Using this index for the 189 dust storms impacting the Phoenix metropolitan area for the period 2010–23, the most severe dust storm occurred on 5 July 2011 and is classified as a “Category 5, Widespread, Long-Duration High-Gust” event. This initial attempt at defining the PHX-DUST scale holds valuable applications for research, education, and applied analyses. It demonstrates that a consortium of interested groups can effectively work to create products benefiting the weather community and general public. Here, the PHX-DUST scale can also serve as a foundational framework for application in other regions where dust storms pose a threat to the well-being of populations, infrastructure, and ecosystems.

Atmosphere

CMIP7 data request: land and land ice priorities and opportunities

The Land and Land Ice Theme in the Coupled Model Intercomparison Project Phase 7 (CMIP7) represents the current understanding of physical processes in land surface ecosystems, hydrology, cryosphere, and their physical interactions with other Earth system components. Simulations from Earth system models (ESMs) could provide crucial information for assessing planetary safety, such as critical tipping elements, and be used to inform climate risks for improving climate impact assessments and policy decisions. This paper presents a collaborative effort to identify scientific opportunities in the Land and Land Ice Theme of the CMIP7 Data Request. The proposed opportunities build upon advances in ESMs, including new freshwater system and land ice processes being included in CMIP7, as well as the scientific community's demand for high-frequency and sub-grid-scale land surface outputs. In total, 25 variable groups that contain 716 variables have been identified to be potentially available to the broad scientific audience for performing analysis in land–atmosphere coupling, hydrological processes and freshwater systems, glacier and ice sheet mass balance and their influence on the sea levels, land use, and plant phenology. Key reflections from this data request effort include advocacy for closer engagement between the user community and modeling groups, reduction in the technical barriers to tracking existing parameters and defining new variables, and more streamlined variable management. These will be essential to enhance the usability and reliability of CMIP7 outputs for climate and Earth system research and applications to a broad audience that relies on the CMIP7 endeavor.

Li, Yue [Univ. of California, Los Angeles, CA (Uni

Lessons Learned from AskGDR: Usage and Impact Analysis of the Geothermal Data Repository's AI Research Assistant: Preprint

In October of 2024, the Department of Energy's (DOE) Geothermal Data Repository (GDR) team officially launched AskGDR, an AI research assistant resulting from the integration of a Large Language Model (LLM) with the metadata and supporting documents associated with GDR datasets. AskGDR allows GDR users to ask deeper questions about the origin of datasets, the methods used to collect them, and the findings they help support. Using Retrieval Augmented Generation (RAG), AskGDR can be used to summarize findings spread across dozens of papers and technical reports or to extract relevant information describing a single data field. However, generative AI is experimental. The National Renewable Energy Laboratory (NREL) has been collecting metrics on AskGDR and documenting lessons learned during its deployment. This paper will outline the efficacy and impact of AskGDR through analysis of its use, operating costs, number and types of questions asked, and the quality of answers provided.

15 GEOTHERMAL ENERGY

Object Proxy Patterns for Accelerating Distributed Applications

Workflow and serverless frameworks have empowered new approaches to distributed application design by abstracting compute resources. However, their typically limited or one-size-fits-all support for advanced data flow patterns leaves optimization to the application programmer—optimization that becomes more difficult as data become larger. The transparent object proxy, which provides wide-area references that can resolve to data regardless of location, has been demonstrated as an effective low-level building block in such situations. Here we propose three high-level proxy-based programming patterns—distributed futures, streaming, and ownership—that make the power of the proxy pattern usable for more complex and dynamic distributed program structures. We motivate these patterns via careful review of application requirements and describe implementations of each pattern. As a result, we evaluate our implementations through a suite of benchmarks and by applying them in three meaningful scientific applications, in which we demonstrate substantial improvements in runtime, throughput, and memory usage.

Distributed Computing

Sierra/SD– User’s Guide for NasGen - 5.30

NasGen provides a path for migration of structural models from NASTRAN bulk data format (BDF) into both an Exodus mesh file and an ASCII input file for Sierra Structural Dynamics (Salinas) and Solid Mechanics (Presto). Many tools at Sandia National Labs use the Exodus format. NasGen was written specifically for Salinas and Presto but should be usable with a number of these packages.

97 MATHEMATICS AND COMPUTING

ECP libraries and tools: An overview

The Exascale Computing Project (ECP) Software Technology and Co-Design teams addressed the growing complexities in high-performance computing (HPC) by developing scalable software libraries and tools that leverage exascale system capabilities. As we enter the exascale era, the need for reusable, optimized software solutions that can handle the unique challenges posed by these systems becomes increasingly important. The primary challenges the ECP teams faced were to create software libraries and tools that are performant on exascale architectures and portable and usable across diverse hardware platforms. Efforts addressed issues related to concurrent execution, memory management, and the integration of heterogeneous computing resources, such as GPUs from multiple vendors. The ECP’s strategy involved a structured development process encompassing the creation, optimization, and deployment of software in collaboration with industry, academia, and national laboratories. The project was organized into several technical areas: co-design of domain-specific suites with target applications, programming models and runtimes, development tools, mathematical libraries, data and visualization tools, and software ecosystem and delivery mechanisms. ECP has successfully developed a large portfolio of software libraries and tools that demonstrate significant improvements in performance and scalability on exascale systems. These products have been integrated into the Department of Energy’s computing facilities, supporting various scientific applications and ensuring robust performance across different hardware setups. ECP advancements in software development for exascale computing highlight the importance of a collaborative and adaptive approach to handling next-generation HPC systems complexities. The lessons learned emphasize the need for continuous engagement with end-users and vendors, and the importance of maintaining a balance between innovation and practical implementation. Future efforts will focus on ensuring scalability, keeping pace with rapid hardware advancements, and further enhancing the interoperability and usability of the software ecosystem. In conclusion, subsequent articles in this special issue provide in-depth discussions and case studies into specific library and tool efforts.

97 MATHEMATICS AND COMPUTING

AI-assisted detector design for the EIC (AID(2)E)

Artificial Intelligence is poised to transform the design of complex, large-scale detectors like ePIC at the future Electron Ion Collider. Featuring a central detector with additional detecting systems in the far forward and far backward regions, the ePIC experiment incorporates numerous design parameters and objectives, including performance, physics reach, and cost, constrained by mechanical and geometric limits. This project aims to develop a scalable, distributed AI-assisted detector design for the EIC (AID(2)E), employing state-of-the-art multiobjective optimization to tackle complex designs. Supported by the ePIC software stack and using G EANT 4 simulations, our approach benefits from transparent parameterization and advanced AI features. The workflow leverages the PanDA and iDDS systems, used in major experiments such as ATLAS at CERN LHC, the Rubin Observatory, and sPHENIX at RHIC, to manage the compute intensive demands of ePIC detector simulations. Tailored enhancements to the PanDA system focus on usability, scalability, automation, and monitoring. Ultimately, this project aims to establish a robust design capability, apply a distributed AI-assisted workflow to the ePIC detector, and extend its applications to the design of the second detector (Detector-2) in the EIC, as well as to calibration and alignment tasks. Additionally, we are developing advanced data science tools to efficiently navigate the complex, multidimensional trade-offs identified through this optimization process.

97 MATHEMATICS AND COMPUTING