Search NASASearch

SEARCH · Search NASA

Results for “Heterogeneous Computing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 127 records · Page 7

BEAST DB: Grand-Canonical Database of Electrocatalyst Properties

We present BEAST DB, an open-source database comprised of ab initio electrochemical data computed using grand-canonical density functional theory in implicit solvent at consistent calculation parameters. The database contains over 20,000 surface calculations and covers a broad set of heterogeneous catalyst materials and electrochemical reactions. Calculations were performed at self-consistent fixed potential as well as constant charge to facilitate comparisons to the computational hydrogen electrode. This article presents common use cases of the database to rationalize trends in catalyst activity, screen catalyst material spaces, understand elementary mechanistic steps, analyze the electronic structure, and train machine learning models to predict higher fidelity properties. Users can interact graphically with the database by querying for individual calculations to gain a granular understanding of reaction steps or by querying for an entire reaction pathway on a given material using an interactive reaction pathway tool. BEAST DB will be periodically updated, with planned future updates to include advanced electronic structure data, surface speciation studies, and greater reaction coverage.

database

Data-driven analysis to understand GPU hardware resource usage of optimizations

With heterogeneous systems, the number of GPUs per chip increases to provide computational capabilities for solving science at a nanoscopic scale. However, low utilization for single GPUs defies the need to invest more money in expensive accelerators. Although related work develops optimizations to improve application performance, none studies how these optimizations impact hardware resource usage or average GPU utilization. Here, this paper takes a data-driven analysis approach in addressing this gap by (1) characterizing how hardware resource usage affects device utilization, execution time, or both, (2) presenting a multiobjective metric to identify important application-device interactions that can be optimized to improve device utilization and application performance jointly, (3) studying hardware resource usage behaviors of several optimizations for a benchmark application, and finally (4) identifying optimization opportunities for several scientific proxy applications based on their hardware resource usage behaviors. Furthermore, we demonstrate the applicability of our methodology by applying the identified optimizations to a proxy application, which improves the execution time, device utilization, and power consumption by up to 29.6%, 5.3% and 26.5% respectively.

Computer science

Future Generation High Performance Computing Center (FG-HPCC): RFI Technical Considerations

Lawrence Livermore National Security, LLC (LLNS) is interested in receiving information about technologies that could be available in the 2029-2030 timeframe that may serve to enable the vision for a Future Generation High Performance Computing (HPC) Center (FG-HPCC) described in this document. The future HPC Center vision has been conceived to meet the future mission needs of the Advanced Simulation and Computing (ASC) Program within the National Nuclear Security Administration (NNSA). LLNS envisions a center composed not of many independent clusters, but of heterogeneous elements accessible to users as a single system. The capabilities will be integrated to create a scalable, flexible, yet tightly coupled computing center capable of integrated HPC, AI, and cloud-like workloads.

97 MATHEMATICS AND COMPUTING

Implementation of Detailed Polyethylene Pyrolysis Kinetics into CFD Simulations using Machine Learning

Municipal solid waste (MSW) and waste plastics have received significant attention due to the issues of waste generation and storage, as well as their potential as an energy resource. High-density polyethylene (HDPE) makes up a large portion of plastic waste and has been the subject of several conversion studies. However, the mechanisms associated with converting HDPE through pyrolysis and gasification are extensive and complex making them difficult to implement into high-fidelity computational fluid dynamic (CFD) simulations. For this project, a primary pyrolysis mechanism containing 42 unique species and 737 heterogeneous reactions was used to generate kinetic data over a range of operating conditions. A machine learning (ML) model was developed to replicate the results of the detailed pyrolysis mechanism while significantly increasing the computational efficiency. A deep operator network (DeepONet) architecture was adopted to train the model using time steps relevant to CFD simulations. The ML used physics-based loss functions to ensure mass conservation. The ML model has been deployed in simple MFiX CFD simulations, single particle, and an experimental drop tube reactor, and has shown promising performance compared to the original scheme.

Houston, Ross

Material Needs and Measurement Challenges for Advanced Semiconductor Packaging: Understanding the Soft Side of Science

This Perspective builds upon insights from the National Institute of Standards and Technology (NIST)-organized workshop, “Materials and Metrology Needs for Advanced Semiconductor Packaging Strategies,” held at the 35th annual Electronics Packaging Symposium in Binghamton, NY, on September 5, 2024. It outlines critical challenges and opportunities related to polymer-based “soft” materials in advanced semiconductor packaging, with emphasis on polymer science, measurement science (metrology), and the strategic development of Research-Grade Test Materials (RGTMs). These efforts, led by the NIST CHIPS team, aim to advance the fundamental understanding of structure-property-processing relationships, promote standardized guidelines and innovative methods for material characterization, and accelerate the development, qualification, and adoption of next-generation packaging materials. The Perspective also distills key insights from the panel discussion with industry experts, emphasizing the need for close collaboration among materials scientists, process engineers, and metrology experts to enable a holistic strategy, further highlighting the importance of cross-sector partnerships among industry, academia, and government to address pressing challenges in packaging materials and processes.

97 MATHEMATICS AND COMPUTING

ORCHA: A performance portability system for extreme heterogeneity

Heterogeneity is the prevalent trend in the rapidly evolving high-performance computing (HPC) landscape in both hardware and application software. The diversity in hardware platforms, currently comprising various accelerators and a future possibility of specializable chiplets, poses a significant challenge for scientific software developers aiming to harness optimal performance across different computing platforms while maintaining the quality of solutions when their applications are simultaneously growing more complex. Code synthesis and code generation can provide mechanisms to mitigate this challenge. We have developed a divide and conquer approach where different aspects of performance are handled by different stand-alone tools that are interfaced with the application through generated code. This portability system, ORCHA, enables users to configure and orchestrate their computations among available resources on a platform by specifying a high-level recipe, thereby permitting a many-to-many paradigm where each recipe results in a different variant of the application. The core design goal is to let users decide the application’s hardware mapping and orchestration by editing only the high-level recipe—without modifying the maintained source code or binding the application to a particular runtime system. Tools in ORCHA distribution are: CG-Kit for translating the recipe into an execution graph; Milhoja to execute the graph by orchestrating data and task movement among hardware resources; and Macroprocessor that enables users to define their own code-shorthand for higher composability and easier management of code variants. Additionally, the design of ORCHA permits tools to work in a plug-and-play mode where the application can build and run without CG-Kit and Milhoja, and either tool can be swapped out for other tools with similar capabilities by modifying the code generation portion of ORCHA. In this paper, we describe the design of ORCHA and the role that code-generation plays in isolating applications from tools. We demonstrate the breadth of configurations ORCHA enables with a case study in which an application configuration is realized on three distinct hardware mappings—a GPU-centric, a CPU/GPU balanced, and a CPU/GPU concurrent layouts by using different recipes.

Lee, Youngjun

Community Requirements Meta-Analysis: Characterizing Needs and Opportunities for HPDF

This High Performance Data Facility (HPDF) Project is creating a new scientific user facility to provide advanced infrastructure for data-intensive science, supporting the DOE’s Office of Science (SC) community. HPDF’s mission is to enable and accelerate scientific discovery by delivering state-of-the-art data management infrastructure, capabilities, and tools. This meta-analysis examines the needs of the breadth of the SC community, captured in publicly available community reports or mission documents. The meta-analysis identifies and provides initial characterization of fifteen core requirements for the HPDF Project team to consider during the conceptual design phase. The fifteen requirements illustrate how scientific work among SC communities requires modern, seamless user experiences across the ASCR Ecosystem to advance the use of large volumes of heterogeneous data. The scientific community requires support for the missing middle of compute between local and HPC to interactively and collaboratively use growing datasets. Data producers and end users will benefit from enhanced data catalogs and portals that improve data access through advanced search of well curated data. The fifteen requirements are examined here organized across five themes for discussion. Examples in each theme illustrate the array of scientific needs that convey the important role that the fully realized and operational High Performance Data Facility will be able to play as an integral part of the evolving ASCR Ecosystem. Our amalgamated data tables from ESnet reports demonstrate ranges to the volumes of data HPDF must be concerned with, but limitations are inherent to this meta-analysis (see Key Challenges & Limitations). Feedback and validation of these requirements along with additional details and emergent community requirements will be gathered through user research and design activities.

97 MATHEMATICS AND COMPUTING

Verifying infectious disease scenario planning for geographically diverse populations

In the face of the COVID-19 pandemic, the literature saw a spike in publications for epidemic models, and a renewed interest in capturing contact networks and geographic movement of populations. There remains a general lack of consensus in the modeling community around best practices for spatiotemporal epi-modeling, specifically as it pertains to the infection rate formulation and the underlying contact or mixing model. We mathematically verify several common modeling assumptions in the literature, to prove when certain choices can provide consistent results across different geographic resolutions, population densities and patterns, and mixing assumptions. The most common infection rate formulation, a computationally low cost per capita infection rate assumption, fails the consistency tests for heterogeneous populations and gravity-weighting assumptions. Future modeling efforts in spatiotemporal disease modeling should be wary of this limitation, particularly when working with more heterogeneous or sparse populations. Our results provide guidance for testing that a model preserves desirable properties even when model inputs mask potential problems due to symmetry or homogeneity. We also provide a recipe for performing this type of verification, strengthening decision support tools.

59 BASIC BIOLOGICAL SCIENCES

RANGE: A robust adaptive nature-inspired global explorer of potential energy surfaces

With the growing demand for realistic representations of chemical structures and the advent of exascale computing, the intelligent sampling of potential energy surfaces and efficient identification of global minima have become more essential but also more feasible. Building on prior studies demonstrating the efficiency of the Artificial Bee Colony (ABC) swarm intelligence algorithm, we report a hybrid metaheuristic framework that integrates the adaptive exploration capabilities of ABC coupled with the exploitation strengths of genetic algorithms (GA) in a scalable, Python-based implementation. The resulting tool, RANGE (Robust Adaptive Nature-inspired Global Explorer), provides seamless interfaces to multiple potential energy evaluators, either directly or via widely used Python libraries, and is designed for high-performance computing environments. We describe the implementation details of RANGE and evaluate its performance, relative to ABC- or GA-alone based algorithms, on a variety of chemical systems, including molecular clusters and heterogeneous surfaces. In conclusion, our results demonstrate RANGE’s efficiency, robustness, and broad applicability in addressing challenging global optimization problems in computational chemistry and materials science.

Algorithms and data structure

A Boundary Element Model for Assessing Large‐Scale Pressurization in Faulted Geological Storage Systems

Assessing large-scale pressurization at the regional scale—a possible outcome of large subsurface storage applications such as wastewater injection and geological carbon sequestration—presents significant computational challenges. These challenges are particularly pronounced when accounting for complex geologic structures with multiple reservoir and caprock layers, fault zones, and wells. This study introduces a computationally efficient model that integrates single-phase semi-analytical solutions with a boundary element (BE) approach. The model simulates pressure propagation in multilayered 3D systems, including vertical faults, caprock, basement, and confining units. We apply this new model to a representative scenario involving CO 2 injection near a partially sealing fault with verification against an independent two-phase flow model. Results demonstrate that our model accurately captures far-field pressure responses and that, outside the CO 2 plume zone, pressure predictions from single-phase and two-phase models are nearly identical. This supports the use of single-phase models like ours for efficient estimation of far-field pressure changes. Additionally, we demonstrate its effectiveness at a large scale, incorporating multiple wells and faults. With its ability to represent multiple wells, fault zones, and geological heterogeneity, our model is well suited for assessments of basin-scale pressurization. Its computational efficiency also makes it a promising tool for integration with optimization frameworks aimed at designing and managing injection strategies in faulted storage systems.

Cihan, A. [Lawrence Berkeley National Laboratory (

The Development of Catalysts for Upgrading of Pyrolysis Vapor for Refinery Feedstocks and Intermediates (CRADA Final Report)

Catalytic fast pyrolysis (CFP) is a versatile technology platform to convert biomass into fungible hydrocarbon transportation fuels and chemical co-products. Key technical barriers to reaching this goal include increasing the product yields and achieving the desired fuel properties for gasoline, diesel, and jet range fuels or blendstocks that would be suitable for introduction into existing refinery unit operations. Overcoming these barriers will require durable catalysts that are effective at upgrading and stabilizing biomass pyrolysis vapors. Towards these goals, this CRADA leveraged NREL experience as a leader in biomass pyrolysis research and Johnson Matthey's (JM) experience as a leader in the production of advanced catalytic materials. The scope spanned CFP catalyst development, characterization, multi-scale reaction testing, and computational modeling. CRADA benefits to DOE, Participant, and U.S. Taxpayer: Assists laboratory in achieving programmatic scope, Uses the laboratory’s core competencies. The purpose of this CRADA was to develop and deploy catalysts for biomass CFP to help achieve cost-competitive biofuels and bio-based products. This was accomplished through a close collaboration between biomass conversion researchers at NREL and catalyst development researchers at JM. Summary of Research Results: Focus Area 1. Foundational research on catalytic conversion and deactivation: Key interactions between pyrolysis vapors and heterogeneous catalysts were probed through catalyst characterization, model compound reaction testing, and atomistic-scale computational modeling. Catalyst development focused on multifunctional materials, which include zeolites, oxides, carbides, and nitrides. Computational modeling identified reaction mechanisms and elucidated surface chemistry to test hypotheses regarding mechanisms of deoxygenation, coupling, cracking, dehydration, coke formation, hydrogen transfer, and aromatic ring reactions. This information was used to design multifunctional catalysts to increase product yields, control product selectivity, and reduce deactivation during CFP and downstream processing steps. The results served to increase fundamental understanding of key catalyst attributes and durability features for the upgrading of biomass pyrolysis vapors. Model compound experiments confirmed the importance of metal-acid bifunctionality for the deoxygenation of lignin-derived phenolic species under hydrodeoxygenation conditions. This insight led to the development of catalysts such as Pt/TiO2 and Mo2C, which were confirmed as high-performing materials during subsequent bench-scale experiments using biomass-derived pyrolysis vapors. This focus area also led to the identification of important catalyst deactivation mechanisms associated with the deposition of inorganic contaminants such as potassium. The molecular-level insight from model compound experiments and computational modeling, shown in Figure 1, informed the development of regeneration procedures that have been shown to be effective for restoration of > 90% of initial catalyst activity. This understanding has subsequently been translated to other catalyst systems, including zeolite materials that can be operated without requirements for co-fed hydrogen.

09 BIOMASS FUELS

Correcting implicit solvation at metal/water interfaces through the incorporation of competitive water adsorption

Conventional continuum solvation models are ubiquitous in computational catalysis, including for describing metal/water interfaces, which are relevant to both solution-phase heterogeneous catalysis and electrocatalysis. Nonetheless, we find that such continuum models qualitatively fail to describe both the adsorption free energy and conformational preference for many organic molecules at such interfaces, largely due to the failure of continuum models to incorporate the role of competitive water adsorption. We develop a simple phenomenological model that accounts for competitive water adsorption and show that the model, when used in conjunction with continuum solvation, provides a dramatic improvement in the description of both adsorption and conformational preference. The model is also extended to additionally incorporate the influence of applied potential at the electrode surface, thus facilitating computationally efficient applications to scenarios including electrocatalysis.

Chemistry

Enabling Innovative Analysis on Heterogeneous Clusters through HTCdaskgateway

High energy particle (HEP) physics research is going through fundamental changes as we move to collect larger amounts of data from the Large Hadron Collider (LHC). Analysis facilities and distributed computing, through HTCs, have come together to create the next pythonic generation of analysis by utilizing HTCdaskgateway, a Dask gateway extension, allowing users to spawn workers compatible with both their analysis and heterogeneous clusters in line with authentication requirements. This is enabling physicists to engage with scientific python in ways they had not before because of domain specific C++ tools. An example of HTCdaskgateway’s use is Fermilab’s Elastic Analysis Facility.

Chavez, Elise [U. Wisconsin, Madison (main)]

Employing artificial intelligence to steer exascale workflows with colmena

Computational workflows are a common class of application on supercomputers, yet the loosely coupled and heterogeneous nature of workflows often fails to take full advantage of their capabilities. We created Colmena to leverage the massive parallelism of a supercomputer by using Artificial Intelligence (AI) to learn from and adapt a workflow as it executes. Colmena allows scientists to define how their application should respond to events (e.g., task completion) as a series of cooperative agents. In this paper, we describe the design of Colmena, the challenges we overcame while deploying applications on exascale systems, and the science workflows we have enhanced through interweaving AI. The scaling challenges we discuss include developing steering strategies that maximize node utilization, introducing data fabrics that reduce communication overhead of data-intensive tasks, and implementing workflow tasks that cache costly operations between invocations. These innovations coupled with a variety of application patterns accessible through our agent-based steering model have enabled science advances in chemistry, biophysics, and materials science using different types of AI. In conclusion, our vision is that Colmena will spur creative solutions that harness AI across many domains of scientific computing.

Workflows

SIREN: Scaling Ion-Traps by REquiring iNnovative Heterogenous Integration

The SIREN (Scaling Ion Traps by Requiring iNnovative heterogenous integration) project explores the feasibility of heterogeneous integration (HI) as a transformative approach to scaling ion traps, a critical technology for advancing quantum computers and atomic clocks. Traditional ion trap architectures face significant challenges in scalability due to limitations in optical access, fabrication techniques, and material constraints. SIREN addresses these challenges by leveraging HI, which combines different materials and fabrication processes to create more complex and efficient ion trap structures. HI integrated structures can be manufactured without compromising the process to maintain compatibility to ion traps. This project focuses on integrating a separately fabricated waveguide with a fully functional ion trap. The respective alignment between the pieces needs to be accurate to less than 2 µm to ensure that the light from the waveguide can overlap with the trapping region. The fine alignment must also be maintained through an ultra-high vacuum bake, a critical step in preparing an ion trap experiment. The project's outcomes suggest that heterogeneous integration is a promising pathway for overcoming current scalability barriers, paving the way for the next generation of quantum technologies. SIREN's findings contribute significantly to the field, offering a scalable solution that could accelerate the development of practical quantum computers and highly accurate atomic clocks.

42 ENGINEERING

Integrating Ultra-Coarse-Grained Protein Models into Accessible Workflows for Multiscale Molecular Dynamics

To capture protein conformational transitions using molecular dynamics (MD), several simulation resolutions covering different spatial and temporal scales are typically needed. All-atom (AA) simulations provide fine resolution, but are computationally infeasible for large systems over longer durations. Coarse-grained (CG) and ultra-coarse-grained (UCG) models have a lower resolution and computational cost while still being able to conserve essential protein features. Prior work on a Multiscale Machinelearned Modeling Infrastructure (MuMMI) combined both AA and CG simulations to study RAS-RAF protein interactions, leveraging CG models for longer time scales and using AA to investigate unusual conformations in greater detail. However, MuMMI is still resource-intensive, and this study aims to maximize exploration of the protein conformational space while reducing computational cost. In this paper, we build on prior work that integrates UCG models based on heterogeneous elastic network modeling (hENM) into the MuMMI workflow. We demonstrate that UCG models enable accurate sampling of protein conformations, focusing on simulating RAS-RAF protein interactions. Using higher-resolution CG Martini simulation data, we can automatically refine intramolecular interactions in UCG models. We present a scalable Python package that uses fluctuations observed in higher-resolution CG Martini simulations to estimate bond coefficients of the UCG model. We built novel machine learning-based backmapping methods to recover more detailed CG Martini structures from UCG structures, using diffusion models to learn the mapping between scales. Finally, we present UCG-mini-MuMMI, an accessible and less compute-intensive version of MuMMI as a resource for the scientific community. Incorporating UCG models into MD studies is applicable to a broad range of systems and proteins, and our study offers insights into the advantages and limitations of these methods.

Chemical structure

Introducing SpaceNet 9 - Cross-Modal Satellite Imagery Registration for Natural Disaster Responses

Computer vision algorithms are increasingly leveraged to accelerate geospatial analysis for disaster response and recovery. As the diversity of remote sensing imagery grows with optical, SAR, and other modalities, a perquisite for analytics is cross-modal image registration. There is a high potential to harness computer vision for this pre-processing requirement toward enabling downstream analytics such as heterogeneous change detection, automated feature extraction, and data fusion. Advancement in these areas has the potential to simplify data wrangling tasks and further accelerate disaster response timelines. The SpaceNet 9 challenge (launching in mid-2024) focuses on addressing the cross-modal image registration problem and demonstrating the utility of such modules on earthquake impacted scenarios. This paper describes the motivation for the SpaceNet 9 and provides a first overview of the dataset, the baseline algorithm, and implications for seeking cross-modal image registration in Earth observation. Code is available at https://github.com/SpaceNetChallenge/SpaceNet9.

Hansch, Ronny

Nanotomography for Quantitative 3D Particle Reconstruction

Particulates are ubiquitous across fuel cycle operations and carry critical information about particle formation, processing, and potential proliferation-related activities. Traditional analytical techniques, including micro-Raman spectroscopy and standard electron microscopy, are often limited in spatial resolution or dimensionality, particularly when used to examine metallic or submicron-scale features. Understanding particle morphology, phase distribution, and internal porosity is essential for constraining formation conditions, thermodynamic environments, and material transport behavior. In this report, we demonstrate the application of plasma focused ion beam nanotomography to reconstruct micron-scale particulates at nanoscale resolution. Using high-resolution backscattered electron imaging and Avizo software, we obtained 3D reconstructions that enabled quantitative analysis of particle morphology, phase composition, and internal voids. Representative examples include a Ta particle with a large central void and a composite particle with embedded tetrahedral crystalline structures. These reconstructions reveal structural and compositional details that are inaccessible through conventional 2D imaging. The results demonstrate that nanotomography provides both qualitative and quantitative insights into particle formation and behavior. Using nanotomography, porosity and phase distributions can be quantified to inform models of particle density, transport, and solidification conditions. Beyond technical insights, the workflow developed here establishes a transferable capability for analyzing heterogeneous particles and has potential applications in bulk materials studies via x-ray computed tomography or other volumetric imaging modalities. Ongoing efforts are focused on optimizing the workflow to process multiple particles simultaneously, increasing throughput and statistical robustness. Overall, this work illustrates the power of nanotomography as a tool for connecting particulate morphology to formation mechanisms, composition, and transport, thereby strengthening analytical capabilities for nuclear forensics, fuel cycle analysis, and related scientific investigations.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS