Search NASA⌕ Search

SEARCH · Search NASA

Results for “computing frameworks”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 361 records · Page 20

Inverse design of hypoeutectoid pearlite steel microstructures using a deep learning and genetic algorithm optimization framework

Goal-oriented microstructure design in metallic materials is a challenging task due to complex structure-property relationships. Traditional experimental and computational approaches are time-intensive and economically inefficient, limiting their applicability for large-scale design space exploration. Here, in this work, we propose an end-to-end framework that integrates deep learning models with genetic optimization to design microstructures with targeted mechanical properties. Deep learning models enable accurate forward design, while their integration with genetic optimization enables efficient inverse design within a few hours, compared to days or weeks using conventional finite element simulations. The framework combines experimental characterization and finite element modeling to analyze the influence of microstructural features on the mechanical behavior of hypoeutectoid steels. Data from both experiments and simulations are used to train the deep learning models. To demonstrate its effectiveness, we apply the framework to 0.63% carbon steel with proeutectoid ferrite and pearlite phases, commonly used in industrial applications. In this study, 2D microstructures were used for modeling, selected primarily for computational efficiency and to establish proof of concept. The framework successfully optimizes microstructures for targeted yield strength, ultimate strength, and stress concentration factors while significantly reducing computational time. Beyond hypoeutectoid steels, this scalable framework can be extended to other material systems and integrated with additive manufacturing, offering an efficient approach for accelerating microstructure design for specific engineering applications.

ConvLSTM↗

Focused Ion Beam Tomography of Alloy 617 Corroded in Molten Chloride Salt

Materials qualification of reactor structural materials is a critical step in rapid implementation of advanced nuclear reactor technologies, particularly to assess the corrosion performance in these designs. Accelerated qualification of reactor structural materials requires incorporating powerful computational toolsets, such as phase field modelling in the Multiphysics Object-Oriented Simulation Environment (MOOSE) framework, to predict the evolution of structural materials due to corrosion. Accordingly, computational toolsets will require experimental data generated at appropriate length scales to validate accuracy. Focused ion beam (FIB) provides a high degree of control over manipulation of materials for analytical purposes, including capturing data on the evolution in the microstructure and elemental composition of materials at the mesoscale, an appropriate length scale for phase field modelling of intergranular diffusion phenomena using the MOOSE framework. For instance, the FEI Helios G4 UX dual beam plasma FIB microscope at the Irradiated Materials Characterization Laboratory (IMCL) is capable of backscatter diffraction (EBSD) and energy-dispersive x-ray spectroscopy (EDS) documenting the evolution in the microstructure and elemental composition, respectively. The Helios can perform EDS and EBSD three-dimensionally (3D) using tomography, which is then combined using different software packages to visualize 3D volumes correlating elemental composition to microstructural data. The purpose of this investigation was to develop a streamlined characterization and data processing workflow for 3D tomography studies on the FEI Helios G4 plasma FIB. The investigation is segmented into three parts: 1) Optimizing the data collection workflow, 2) identifying appropriate data processing and visualization software (i.e. DREAM.3D, MIPAR, and VGStudioMax), and 3) establishing an infrastructure for public release. The optimization of the data collection workflow is in collaboration with members of the U220 department to setup formal training on the tomography operation of the G4, through ThermoFisher Scientific, and exploring DREAM.3D, MIPAR, and VGStudioMax data processing/visualization software packages. VGStudioMax currently demonstrates the most promise for future use. Optimization of the data collection and processing workflow is still ongoing. A collaboration with INL High Performance Computing (HPC) established an open-source license for expediting the public release of FIB tomography datasets through HPC. FIB tomography data generated by the G4 will provide comprehensive data for validating 3D phase field mesoscale modelling tools within the MOOSE framework for accelerated qualification of reactor structural materials.

Copeland-Johnson, Trishelle↗

Operator learning for energy-efficient building ventilation control with computational fluid dynamics simulation of a real-world classroom

Energy-efficient ventilation control plays an important role in reducing building energy consumption while ensuring occupant health and comfort. While Computational Fluid Dynamics (CFD) simulations provide detailed and physically accurate representations of indoor airflow, their high computational cost limits their use in real-time building control. In this work, we present a neural operator learning framework that combines the physical accuracy of CFD with the computational efficiency of machine learning to enable building ventilation control with the high-fidelity fluid dynamics models. Our method jointly optimizes the airflow supply rates and vent angles to reduce energy use and adhere to air quality constraints. We train an ensemble of neural operator transformer models to learn the mapping from building control actions to airflow fields using high-resolution CFD data. This learned neural operator is then embedded in an optimization-based control framework for building ventilation control. Experimental results show that our approach achieves significant energy savings compared to maximum airflow rate control, rule-based control, as well as data-driven control methods using spatially averaged CO 2 prediction and deep learning–based reduced-order models, while consistently maintaining safe indoor air quality. These results highlight the practicality and scalability of our method in maintaining energy efficiency and indoor air quality in real-world buildings.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Control of Permanent Porosity in Type 3 Porous Liquids via Solvent Clustering

Porous liquids (PLs) are an exciting new class of materials for carbon capture due to their high gas adsorption capacity and ease of industrial implementation. They are composed of sorbent particles suspended in a nonadsorbed solvent, forming a liquid with permanent porosity. While PLs have a vast number of potential compositions based on the number of solvents and sorbent materials available, most of the research has been focused on the selection of the sorbent rather than the solvent. Therefore, PL design criteria on the supramolecular structures of the solvent are explored to create a fundamental understanding of how the solvent enables PL formation for rapid discovery of new PL compositions. Atomistic molecular dynamics simulation of eight solvents with a range of molecular sizes, shapes, and intramolecular bonding was performed, identifying that the shape and size of molecular clusters formed in the solvent are the driving predictor of PL formation rather than the size of the individual solvent molecule. The results demonstrate a significant departure from common approaches to PL formation based on the steric exclusion of solvent molecules from the sorbent via the size of the pore aperture. A modeling and experimental validation study further supports these findings. In conclusion, through this computational material design study, a previously unexplored mechanism in PL formation, solvent–solvent clustering, is identified as a critical factor for the accelerated discovery of liquid phase carbon capture materials.

Carbon capture↗

Improving Spectral Resolution from Real-time Evolution for Correlated Systems

Abstract The quality of numerically simulated spectra using real-time evolution methods for strongly correlated systems is affected by both the length of simulation time and the system size, limiting resolution in both frequency and momentum. In this work, we propose a computationally cheap, linear autoregressive machine learning-based framework to extend short-time and short-distance results over a wider range. We use the proposed method to extend the lesser Green’s function for both the Hubbard model and the much more computationally challenging Hubbard-extended Holstein model. This technique significantly improves both the frequency and momentum resolution of the single-particle removal spectrum $${\mathcal{A}}(k,\omega )$$ A ( k , ω ) , allowing the observation of otherwise obscured spectral features due to electron-phonon coupling.

Tang, Ta↗

A Contextually-Aware Sensitivity Analysis to Guide the Design of Randomized Least Squares Solvers in Applications

Our work on the DOE-sponsored project “A Contextually-Aware Sensitivity Analysis to Guide the Design of Randomized Least Squares Solvers in Applications,” was an effort to address critical challenges in nu merical computing and its applications to optimization. The increasing demand for robust and scalable solutions to large-scale linear algebra problems has highlighted the limitations of traditional approaches, particularly in heterogeneous and extreme-scale computing environments. Randomized Numerical Linear Algebra (RandNLA) offers a promising framework to address these challenges, and this proposal builds on this foundation by introducing innovations in sensitivity analysis and computational adaptability.

97 MATHEMATICS AND COMPUTING↗

pH-Dependent Vibrational Dynamics Drives Excited-State Quenching in the Phycobiliprotein Complex PC645

Phycocyanin 645 (PC645) is a closed-form lightharvesting complex found in the lumen of the photosynthetic membrane of cryptophyte algae. These peripheral antenna complexes contain bilin chromophores that absorb sunlight and transfer excitation energy to the core antenna complexes embedded in the thylakoid membrane. The location of cryptophyte antenna complex on the luminal side of the membrane is unusual. During photosynthetic activity, the pH of the lumen drops, by up to two pH units. There is little known about how this pH-change affects the light-harvesting complexes. In this study, we report multiscale simulations using a computationally efficient density functional tight-binding framework to investigate the spectroscopy and excitation energy transfer in the PC645 complex. Complementary experiments were conducted using both steady-state and time-resolved spectroscopic measurements at low, neutral, and high pH values. Our study shows that (de)protonation of specific bilin pigments, namely, the mesobiliverdins (MBVs), modulates the excitation energies, excitonic couplings, and spectral densities. These changes cause excitation transfer rates to increase by up to a factor of two to three, leading to pH-dependent energy transfer pathways in the complex. Using this model, we calculated the pH-dependent fluorescence quantum yield of the system, obtaining quantitative agreement with the experimental results. These computational simulations, supported by experiments, identify MBVs as a more prominent excitation sink than previously realized, and that this role is tuned by pH.

Maity, Sayan [Constructor Univ., Bremen (Germany);↗

Computational Analysis of the Energetic Stability of High-Entropy Structures of a Prototypical Lanthanide-Based Metal–Organic Framework

High-entropy materials are characterized by their complex compositions, typically comprising five or more elements in near-equiatomic proportions. Applying this concept to metal ions in metal−organic frameworks (MOFs) has paved the way for exploring a new class of high-entropy MOFs. While the compositional strategy of high-entropy materials leverages configurational entropy to aid thermodynamic stability, it also poses significant analytical challenges due to the vast compositional landscape and diverse phases that these materials can adopt. We present a computational study of several complexities associated with selecting potential high-entropy versions of a prototype lanthanidebased MOF. We compute the energetics of metal mixing of these heterometallic MOFs using density functional theory (DFT) and machine learning interatomic potential (MLIP) methods. The use of MLIP methods allows a systematic exploration of the convex hull of thermodynamically stable MOF structures containing up to 5 distinct metals.

Chemical structure↗

An AI-Based 3D Bat Movement Tracking System at Wind Energy Facilities Using Multi-Thermal Video Cameras

The poster at the 15th Wind Wildlife Research Meeting discusses how to leverage the potential of real-time thermal-imaging methodologies in quantifying nocturnal bat activities at wind turbines, using 3D computer vision techniques within a deep learning framework. This innovation enables the automatic detection and classification of bats, birds, and insects in thermal-imaging videos captured at wind turbine sites, facilitating efficient and accurate data analysis for enhanced understanding and mitigation of bat-wind turbine interactions.

AI↗

DTLMod: A simulation framework for in situ workflow optimization

In situ processing workflows have become essential for coping with the explosion in data volume and velocity in large-scale scientific computing, providing domain scientists with early insights at runtime. Multiple frameworks implement this paradigm through a data transport layer (DTL), offering different data access modes and deployment schemes, but researchers currently lack the appropriate tools to assess design and deployment options before committing to costly real experiments. We introduce DTLMod, an open-source simulated DTL that enables performance evaluation of in situ workflow configurations at scale. Built on SimGrid, it links into any SimGrid-based simulator and is available in C++ and Python. We evaluate DTLMod along four axes: scalability (tens of thousands of simulated processes across interconnected clusters in seconds, with linear memory scaling), versatility (three implementation variants trading fidelity for speed), accuracy (simulated times faithfully reflecting real behavior), and practical utility (two use cases demonstrating evidence-based workflow design decisions).

Suter, Fred [ORNL] (ORCID:0000000319021955)↗

Transitioning GlideinWMS, a multi domain distributed workload manager, from GSI proxies to tokens and other granular credentials

GlideinWMS is a distributed workload manager that has been used in production for many years to provision resources for experiments like CERN’s CMS, many Neutrino experiments, and the OSG. Its security model was based mainly on GSI (Grid Security Infrastructure), using X.509 certificate proxies and VOMS (Virtual Organization Membership Service) extensions. Even when other credentials, like SSH keys, were used to authenticate with resources, proxies were also added all the time, to establish the identity of the requestor and the associated memberships or privileges. This single credential was used for everything and was, often implicitly, forwarded wherever needed. The addition of identity and access tokens and the phase-out of GSI forced us to reconsider the security model of GlideinWMS, to handle multiple credentials which can differ in type, technology, and functionality. Both identity tokens and access tokens are supported. GSI proxies even if no more mandatory, are still used, together with various JWT (JSON Web Token) based tokens and other certificates. The functionality of the credentials, defined by issuer, audience, and scope, also differ: a credential can allow access to a computing resource, or can protect the GlideinWMS framework from tampering, or can grant read or write access to storage, can provide an identity for accounting or auditing, or can provide a combination of any the formers. Furthermore, the tools in use do not include automatic forwarding and renewal of the new credentials so credential lifetime and renewal requirements became part of the discussion as well. In this paper, we will present how GlideinWMS was able to change its design and code to respond to all these changes.

97 MATHEMATICS AND COMPUTING↗

Enforcing global constraints for the dispersion closure problem: τ 2 -SIMPLE algorithm

Permeability and effective dispersion tensors are critical parameters to characterize flow and transport in porous media at the continuum scale. Homogenization theory defines a framework in which such effective properties are first computed from solving a closure problem in a repeating unit cell of the periodic microstructure and then used in a macroscopic formulation for efficient computation. The closure problem is formulated as a local boundary value problem subjected to global constraints, which guarantee the uniqueness of the solution and can be difficult to satisfy for complex geometries and at high flow conditions. These constraints also ensure that pore-scale pressure, velocity, and concentration fields can be accurately reconstructed from the closure variable. Building on a previous work, here we present a framework that allows to satisfy global constraints associated to both the permeability and the dispersion closure problems by introducing two artificial time scales. The algorithm, called τ 2 -SIMPLE, computes both permeability and effective dispersion given an arbitrarily complex geometry and flow condition. Furthermore, this algorithm is demonstrated to be accurate for both 2D and 3D geometries across varying flow conditions, and thus it can be used to quickly characterize effective properties from porous media images in many applications.

97 MATHEMATICS AND COMPUTING↗

Continuous integration data-driven platform of industrial-scale subsurface storage for real-time analytics

This project helped address the growing need for efficient and scalable models to support geological carbon and energy storage, which are crucial for achieving net-zero emissions. Traditionally accurate high-fidelity numerical models have been used to simulate relevant storage processes under a handful of processes, however such models are computationally demanding, making uncertainty quantification impractical. Consequently, we first developed a machine learning framework, based on Graph Neural Operators (GNOs), to improving the accuracy of model predictions for a fixed computational budget. We then developed an Ensemble of Improved Neural Operators (ENO), which uses bagging and Monte Carlo dropout techniques, to further improve prediction accuracy. Lastly, we developed the way to explain progressive transfer learning methods to reduce the amount of training data and computational cost of training (i.e., reduce trainable parameters) when using our models for multiple storage sites. Our numerical investigation, which used real-world case studies, demonstrated that our framework can significantly improve the safety and efficiency of geological storage operations, with potential applications in other domains such as geothermal reservoirs and climate modeling.

54 ENVIRONMENTAL SCIENCES↗

A Benchmark Suite for Evaluating Scientific AI Workloads on GPUs

AI applications have been steadily increasing in the allocation portfolio among leadership computing facilities. These applications depend on deep learning frameworks with hardware acceleration and underlying software systems. With the rapid development of applications, software stacks, and hardware devices, it is essential to evaluate the performance of core operations in AI workloads for direction of optimizations and procurement of next-generation high-performance computing (HPC) infrastructures. Currently, most benchmarks lack scientific AI workloads. So, we present DeepKernelBench and the experimental results of evaluating the benchmark suite for early observations and performance comparisons on datacenter GPUs using representative workloads for scientific AI, including Attentions, General matrix multiplications, Geometrics and Fourier neural operations.

Jin, Zheming [Advanced Micro Devices (AMD)]↗

Exploring Flood Predictability in Taiwan through Coupled Atmospheric–Hydrological and High-Performance Hydrodynamic Models

Effective flood simulation capabilities can tremendously support early warning and disaster prevention. To examine the applicability of a fully physics-based and high-performance flood simulation and forecasting modeling framework for a flood-prone region in Taiwan, we conduct a numerical experiment that couples the Weather Research and Forecasting (WRF) Model, WRF-Hydrological modeling system (WRF-Hydro), and the Two-Dimensional Runoff Inundation Toolkit for Operational Needs (TRITON) to perform integrated rainfall, streamflow, and flood simulations. Furthermore, we first use the coupled WRF and WRF-Hydro (WWH) to predict rainfall and streamflow and then drive TRITON with the predicted streamflow hydrographs to simulate flood depth and inundation area. With the refined spatial resolution and parameterization, this framework can better predict rainfall with reasonable spatial patterns. Although WWH could overestimate the amount of rainfall in some areas, the uncertain rainfall–streamflow predictions produce reasonable flood maps able to pinpoint regions at risk of flooding. In terms of model efficiency, the graphics processing unit–based computation can yield a speed-up factor as high as ∼13 compared to the central processing unit–based computation, promoting the efficacy of the coupled modeling framework in practical real-time flood forecasting.

Coupled models↗

Ambient and Initial Temperature Effects on Energy Consumption Rate Modeled in FASTSim

Ambient and initial temperatures significantly impact the energy consumption rate (ECR) of battery electric vehicles (BEVs) due to auxiliary loads and the temperature dependence of battery efficiency. This study introduces a streamlined, physics-based thermal modeling approach within the FASTSim tool that bridges the gap between oversimplified constant-load models and computationally expensive high-fidelity simulations. By employing a lumped thermal mass framework, the model captures fundamental energy balances and critical non-linear energy penalties while maintaining the computational efficiency required for expansive sensitivity studies. The simulations evaluated a compact BEV hatchback with a resistive heater over city (UDDS) and highway (HWFET) test cycles. Compared to a 22 degrees Celsius initial and ambient temperature baseline, a -7 degrees Celsius initial/ambient temperature resulted in a 221% increase in the ECR for the city cycle and a 100% increase for the highway cycle. Conversely, a 45 degrees Celsius initial / 40 degrees Celsius ambient temperature resulted in a 40% increase for UDDS and an 18% increase for HWFET. These results demonstrate that while cold conditions impose the most severe energy penalties due to resistive heating, the impact is consistently more pronounced in city driving where auxiliary loads represent a larger proportion of total energy. This lightweight yet robust framework enables researchers to rapidly quantify BEV thermal sensitivity across diverse climates without the need for high-overhead simulation environments.

33 ADVANCED PROPULSION SYSTEMS↗

Multi-physics melt pool modeling and process optimization for laser direct energy deposition of Nb-based refractory C103: Defect formation, geometric precision, and process mapping

Recent developments in additive manufacturing (AM) technology have reignited interest in the fabrication of the Nb-based refractory C103 alloy offering solutions to the challenges posed by traditional manufacturing methods. However, the limited numerical and experimental studies on laser direct energy deposition (DED) of C103 have hindered the understanding of the relationships between process parameters and build quality. This has made it challenging to consistently produce parts with the desired quality and microstructure suitable for critical applications. In this study, we focus on optimizing the laser DED process for C103 by employing a hybrid approach that combines experimental techniques and computational fluid dynamics (CFD). This approach facilitates the development of process maps for defect detection and geometric precision. To achieve this, multi-layer C103 samples were fabricated using laser DED under various process parameters, enabling the creation of a process map for defect detection. Additionally, a multi-physics, multiphase simulation framework was developed within a high-performance computing (HPC) environment to establish process maps for geometric precision. Using these process maps, printability windows were identified for achieving both the desired geometric accuracy and defect-free prints. It was observed that prints with a power-to-velocity (P/V) ratio close to unity resulted in defect-free outcomes. This study provides a foundation for reducing design lead time and rejected parts, ultimately optimizing the laser DED process for C103.

Defect formation and geometric precision↗

Practical Implementation of GPU-based Computing at the Grid Edge for Resilience Scenarios

This paper presents a practical implementation of GPU-accelerated computing at the grid edge to enhance power system resilience through next-generation smart meters. Advanced Metering Infrastructure (AMI) systems rely predominantly on centralized processing architectures, which limit real-time response capabilities during grid disturbances. This work proposes the integration of GPU-enabled computational platforms directly within smart meter to enable local execution support for power system analytics, fault detection algorithms, and optimization routines. The proposed framework uses the Julia programming language to leverage highperformance parallel computing capabilities while maintaining code portability and development efficiency. We use two experimental scenarios to benchmark the computational feasibility of this approach: sparse linear system solutions representative of power flow analyses, and multi-stage production cost simulations incorporating unit commitment and economic dispatch operations. Results demonstrate that computationally intensive power system algorithms, such as those supporting resilience scenario calculations, can be effectively executed at the distribution edge using commercially available embedded GPU hardware. Keywords—GPU acceleration, edge computing, smart meters, grid resilience, AMI, resilience.

De Souza, Reubun [School of Electrical Engineering↗