Search NASASearch

SEARCH · Search NASA

Results for “Graphics”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

Portable Acceleration of CMS Computing Workflows with Coprocessors as a Service

Computing demands for large scientific experiments, such as the CMS experiment at the CERN LHC, will increase dramatically in the next decades. To complement the future performance increases of software running on central processing units (CPUs), explorations of coprocessor usage in data processing hold great potential and interest. Coprocessors are a class of computer processors that supplement CPUs, often improving the execution of certain functions due to architectural design choices. We explore the approach of Services for Optimized Network Inference on Coprocessors (SONIC) and study the deployment of this as-a-service approach in large-scale data processing. In the studies, we take a data processing workflow of the CMS experiment and run the main workflow on CPUs, while offloading several machine learning (ML) inference tasks onto either remote or local coprocessors, specifically graphics processing units (GPUs). With experiments performed at Google Cloud, the Purdue Tier-2 computing center, and combinations of the two, we demonstrate the acceleration of these ML algorithms individually on coprocessors and the corresponding throughput improvement for the entire workflow. This approach can be easily generalized to different types of coprocessors and deployed on local CPUs without decreasing the throughput performance. We emphasize that the SONIC approach enables high coprocessor usage and enables the portability to run workflows on different types of coprocessors.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND

Single-atom materials boosting wearable orthogonal uric acid detection

Abstract Uric acid (UA) is a vital biomarker for the diagnosis and management of various health conditions, including cardiovascular diseases, gout, kidney disorders, metabolic syndrome, and wound healing. Despite significant advances in wearable sensor technology, challenges persist in developing wearable sensors that are capable of maintaining high sensitivity, selectivity, and stability. In this study, we present an epidermal sensing platform enhanced with single-atom materials (SAMs) designed for flexible and orthogonal electrochemical detection of UA. We designed and synthesized an SAM with Fe-N 5 active sites to boost the electrochemical sensing signals, integrating it with laser-engraved graphene (LEG) to fabricate a wearable SAM-based UA patch sensor. This design provides superior UA detection performance compared to sensors based on conventional nanomaterials. In addition, we enhanced the detection accuracy and range by using an orthogonal approach that combines direct oxidation through differential pulse voltammetry (DPV) along with parallel biocatalytic amperometric detection. The resulting SAM-based UA orthogonal sensor patch demonstrated exceptional performance in wearable applications through tests measuring sweat UA levels in subjects before and after consuming a purine-rich diet. Graphical Abstract

Ding, Shichao

Energetics of the nucleation and glide of disconnection modes in symmetric tilt grain boundaries

Grain boundaries (GBs) evolve by the nucleation and glide of disconnections, which are dislocations with a step character. In this work, motivated by recent success in predicting GB properties such as the shear coupling factor and mobility from the intrinsic properties of disconnections, we develop a systematic method to calculate the energy barriers for the nucleation and glide of individual disconnection modes under arbitrary driving forces and a quasi-2D setting. This method combines tools from bicrystallography to enumerate disconnection modes and the Nudged elastic band (NEB) method to calculate their energetics, yielding minimum energy paths and atomistic mechanisms for the nucleation and glide of each disconnection mode. We apply the method to accurately predict shear coupling factors of $[001]$ symmetric tilt grain boundaries in Cu. Particular attention is paid to the boundaries where the dislocation-based disconnection nucleation model produces incorrect nucleation barriers. We demonstrate that the method can accurately compute energy barriers and predict shear-coupling factors in the low-temperature regime. For certain disconnection modes in which the assumptions underlying our method do not hold, we report upper bounds on the energy barriers for disconnection nucleation and glide. In addition, the NEB trajectories reveal interesting phenomena such as the dissociation of a higher energy mode into lower energy modes, and in some cases, shear coupling being mediated by partial disconnections, wherein the GB structure temporarily changes to a metastable state before reverting back to its original structure. Graphical abstract

36 MATERIALS SCIENCE

An integrated modeling framework with open architecture for phase field simulation of multi-component alloys

An integrated modeling framework (PanPhaseField) has been developed, which enables a direct and fast coupling between CALPHAD calculations and large-scale phase field simulations for multi-component alloys. Further, it adopts an open architecture allowing for integration of user-defined phase field models in a plug-and-play manner by taking full advantage of the user-friendly graphical interface of Pandat software. The developed modeling platform becomes an enabling tool that can be used to simulate the evolution of spatially varying microstructures of industrial complex alloys for various engineering applications.

36 MATERIALS SCIENCE

INSPIRED: Inelastic neutron scattering prediction for instantaneous results and experimental design

Inelastic neutron scattering (INS) has unique advantages in probing how atoms vibrate and how the vibrations propagate and interact. Such dynamic information is crucial in understanding various material properties, from heat capacity, thermal conductivity, phase transitions, and chemical reactions to more exotic quantum behavior. The analysis and interpretation of the INS spectra often start from a model structure of the sample, followed by a series of calculations to obtain the simulated spectra to compare with experiments. The conventional way to perform such calculations usually requires significant time, computing resources, and specialized expertise. Here, we present a new program named INSPIRED (Inelastic Neutron Scattering Prediction for Instantaneous Results and Experimental Design), which enables users to perform rapid INS simulations in several different ways on their personal computers in just a few clicks, with the crystal structure as the only input file. Specifically, the users can choose a pre-trained symmetry-aware neural network (coupled with an autoencoder) to predict the phonon density of states (DOS), 1D S(E) and 2D S(|Q|,E) spectra for any given structure. One can also choose an existing density functional theory (DFT) calculation from a database (containing over 12,000 crystals), and quickly obtain the simulated INS spectra for single crystals and powders. It is also possible to use pre-trained universal machine learning force fields to relax a given crystal structure, calculate the phonon dispersion and DOS, and, subsequently, the INS spectra. All these functions are implemented with a PyQt graphic user interface. Finally, we expect these new tools will benefit broad user communities and significantly improve the efficiency of experiment design, execution, and data analysis for INS.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS

An immersed interface method for microstructure-scale electrochemical battery models: numerical formulation and performance portable implementation

We present the numerical formulation, verification, and performance portable implementation of an immersed interface method for microstructure scale electrochemical modeling of batteries. The innovation in this approach is the resolution of chemical species and electrostatic potential discontinuities at active interfaces without the use of interface conforming unstructured grids. A unified formulation on Cartesian grids for all domains (electrodes and electrolyte) is used with interfacial flux conditions applied using volume fraction or “color” function gradients. We have developed one dimensional and two dimensional test cases with analytic solutions for electrochemical modeling using which we verified the consistency and accuracy of our scheme. Our solver is also validated against solutions from a macroscale model and an unstructured multi-subdomain solver for a full lithium ion cell. We then demonstrated the utility of our solver on an image-based complex battery electrode microstructure at high charging rate. Our technique also exhibits good scalability on distributed memory architectures using central processing units (CPU), with problem sizes up to 1.8 billion degrees of freedom and with 5400 ranks. Initial performance studies of our open-source performance portable solver showed about 70 times speed up using a graphics processing unit (GPU) compared to single compute core for a problem with 4 million cells.

25 ENERGY STORAGE

Harnessing distributed GPU computing for generalizable graph convolutional networks in power grid reliability assessments

Although machine learning (ML) has emerged as a powerful tool for rapidly assessing grid contingencies, prior studies have largely considered a static grid topology in their analyses. This limits their application, since they need to be re-trained for every new topology. Here, this paper explores the development of generalizable graph convolutional network (GCN) models by pre-training them across a range of grid topologies and contingency types. We found that a GCN model with auto-regressive moving average (ARMA) layers with a line graph representation of the grid offered the best predictive performance in predicting voltage magnitudes (VM) and voltage angles (VA). We introduced the concept of phantom nodes to consider disparate grid topologies with a varying number of nodes and lines. For pre-training the GCN ARMA model across a variety of topologies, distributed graphics processing unit (GPU) computing afforded us significant training scalability. The predictive performance of this model on grid topologies that were part of the training data is substantially better than the direct current (DC) approximation. Although direct application of the pre-trained model to topologies that are not part of the grid is not particularly satisfactory, fine-tuning with small amounts of data from a specific topology of interest significantly improves predictive performance. In general, this paper highlights the feasibility of training large-scale GNN models to assess the reliability of power grids by considering a wide variety of grid topologies and contingency types. With the advent of foundational models in ML and the exponential increase in GPU computing clusters, generalizable ML models will significantly enhance how utilities manage power systems and make decisions in real-time or near-real-time.

24 - POWER TRANSMISSION AND DISTRIBUTION

Generalized fiducial inference on differentiable manifolds

We introduce a novel approach to inference on parameters that take values in a Riemannian manifold embedded in a Euclidean space. Parameter spaces of this form are ubiquitous across many fields, including chemistry, physics, computer graphics, and geology. Here, this new approach uses generalized fiducial inference (GFI) to obtain a posterior-like distribution on the manifold, without needing to know local parameterizations that map to the constrained space from an unconstrained Euclidean space. Using mathematical tools from Riemannian geometry, we construct a constrained generalized fiducial distribution (CGFD). A Bernstein-von Mises-type result for the CGFD, which provides intuition for how the desirable asymptotic qualities of the unconstrained generalized fiducial distribution are inherited by the CGFD, is provided. To illustrate the practical use of the CGFD, we provide a proof-of-concept example in the context of a linear logspline density estimation problem, and demonstrate that CGFD-based confidence sets exhibit desirable coverage properties via simulation. As an application, we fit a CGFD to COVID-19 case count data from North Carolina, USA.

97 MATHEMATICS AND COMPUTING

Accelerating high-order continuum kinetic plasma simulations using multiple GPUs

Kinetic plasma simulations solve the Vlasov-Poisson or Vlasov-Maxwell equations to evolve scalar-variable distribution functions in position-velocity phase space and vector-variable electromagnetic fields in configuration space. The immense computational cost of evolving high-dimensional variables, and their large number of degrees of freedom, often limits the utility of continuum kinetic simulations and presents a challenge when it comes to accurately simulating real-world physical phenomena. To address this challenge, we present techniques that accelerate and minimize the computational work required for a scalable Vlasov-Poisson solver. We show theoretical hardware compute and communication bounds for solving a fourth-order finite-volume Vlasov-Poisson system. These bounds are then used to inform and evaluate the design of performance portable algorithms for a multiple graphics processing unit (GPU) accelerated version of the Vlasov-Poisson solver VCK-CPU [1]. We demonstrate that the multi-GPU Vlasov solver implementation, VCK-GPU, simultaneously minimizes required inter-process data transfer while also being bounded by the machine network performance limits. This results in an overall strong scaling speedup per timestep of up to 40x in three-dimensional phase space (one position, two velocity coordinates) and 54x in four dimensional phase space (two position, two velocity coordinates) and a 341x increase in simulation throughput of the GPU accelerated code over the existing CPU code. The GPU code is also able to weak scale up to 256 compute nodes and 1024 GPUs. In conclusion, we demonstrate that the improved compute performance enables exploring configurations which were previously computationally infeasible, including resolving fine-scale distribution function filamentation and multi-species dynamics with realistic electron-proton mass ratios.

Continuum kinetics

Hardware acceleration for HPS algorithms in two and three dimensions

We provide a flexible, open-source framework for hardware acceleration, namely massively-parallel execution on general-purpose graphics processing units (GPUs), applied to the hierarchical Poincaré–Steklov (HPS) family of algorithms for building fast direct solvers for linear elliptic partial differential equations. To take full advantage of the power of hardware acceleration, we propose two variants of HPS algorithms to improve performance on two- and three-dimensional problems. In the two-dimensional setting, we introduce a novel recomputation strategy that minimizes costly data transfers to and from the GPU; in three dimensions, we modify and extend the adaptive discretization technique of Geldermans and Gillman [1] to greatly reduce peak memory usage. We provide an open-source implementation of these methods written in JAX, a high-level accelerated linear algebra package, which allows for the first integration of a high-order fast direct solver with automatic differentiation tools. We conclude with extensive numerical examples showing our methods are fast and accurate on two- and three-dimensional problems.

Fast direct solvers

Constellation: The autonomous control and data acquisition system for dynamic experimental setups

The operation of instruments and detectors in laboratory or beamline environments presents a complex challenge, requiring stable operation of multiple concurrent devices, often controlled by separate hardware and software solutions. These environments frequently undergo modifications, such as the inclusion of different auxiliary devices depending on the experiment or facility, adding further complexity. The successful management of such dynamic configurations demands a flexible and robust system capable of controlling data acquisition, monitoring experimental setups, enabling seamless reconfiguration, and integrating new devices with limited effort. This paper presents Constellation, a flexible and network-distributed control and data acquisition software framework tailored to laboratory and beamline environments, that addresses the limitations of existing solutions. The framework is designed with a focus on extensibility, providing a streamlined interface for instrument integration. It supports efficient system setup via network discovery mechanisms, promotes stability through autonomous operational features, and provides comprehensive documentation and supporting tools for operators and application developers such as controllers and logging interfaces. At the core of the architectural design is the autonomy of the individual components, called satellites, which can make independent decisions about their operation and communicate these decisions to other components. This paper introduces the design principles and framework architecture of Constellation, presents the available graphical user interfaces, shares insights from initial successful deployments, and provides an outlook on future developments and applications.

Autonomy

Iterative methods in GPU-resident linear solvers for nonlinear constrained optimization

Linear solvers are major computational bottlenecks in a wide range of decision support and optimization computations. The challenges become even more pronounced on heterogeneous hardware, where traditional sparse numerical linear algebra methods are often inefficient. For example, methods for solving ill-conditioned linear systems have relied on conditional branching, which degrades performance on hardware accelerators such as graphical processing units (GPUs). To improve the efficiency of solving ill-conditioned systems, our computational strategy separates computations that are efficient on GPUs from those that need to run on traditional central processing units (CPUs). Our strategy maximizes the reuse of expensive CPU computations. Iterative methods, which thus far have not been broadly used for ill-conditioned linear systems, play an important role in our approach. In particular, we extend ideas from Arioli et al., (2007) to implement iterative refinement using inexact LU factors and flexible generalized minimal residual (FGMRES), with the aim of efficient performance on GPUs. In conclusion, we focus on solutions that are effective within broader application contexts, and discuss how early performance tests could be improved to be more predictive of the performance in a realistic environment.

97 MATHEMATICS AND COMPUTING

A High-Performance Discrete-Element Framework for Simulating Flow and Jamming of Moisture Bearing Biomass Feedstocks

We developed and verified a high-performance open-source discrete element method (DEM) solver with simultaneously-supported feedstock-specific interaction models, including bonded-sphere, liquid bridge, cohesion, and non-linear contact models. Our solver uses parallel data structures on hybrid central and graphics processing unit (CPU/GPU) architectures, with favorable strong scaling performance observed for large problem sizes comprised of (100 M particles), and 4X single-node GPU speedup. The particles for corn stover feedstock were conceptualized and calibrated based on experimental measurements and results. Sensitivity analyses demonstrate that the mass flow rate from a wedge hopper is governed primarily by moisture content, friction coefficient, and cohesion energy density. The model is used to reproduce experimentally observed hopper jamming results, highlighting that the experimental no-flow trends can only be achieved by using non-spherical particles, liquid bridge and cohesion models, highlighting the importance of using concurrent feedstock specialized models for the effective representation of biomass material handling problems.

bioenergy

A Review of the Lawrence Livermore Nuclear Accident Dosimeter 1980s-present

A Nuclear Accident Dosimetry program is a federal requirement for all facilities that have the potential to have a criticality accident. Personnel Nuclear Accident Dosimeter (PNAD) theory and analytical procedures are driven by various scientific needs and interacting regulations. A brief history of the status of USA Department of Energy (DOE) nuclear accident dosimetry regulations, recommendations, and performance testing criteria are given. Then, the history of the Lawrence Livermore National Laboratory (LLNL) PNAD is explored, including changes in the physical dosimeter and adjustments of the analysis method through the last four decades. Finally, the performance of LLNL’s PNAD at criticality accident intercomparison training exercises since 2009 is explored. In general, reported neutron doses have been within or close to DOE-STD-1098 performance criteria while reported gamma doses have been outside of DOE-STD-1098 performance criteria. Reported total absorbed doses have varied in meeting ANSI/HPS N13.3 and ANSI/HPS N13.3 (R2019) performance criteria. Dosimetry staff retirement and turnover have left historical knowledge gaps, yet provided opportunities within the NAD program at LLNL. This review paper serves as an overview of the history and status of the NAD program. Brief technical, procedural and programmatic recommendations to improve LLNL’s NAD program are given. Technical recommendations include investigating orientation factors through modeling or empirical experimentation, investigating gamma dosimetry methods for high-dose scenarios, and exploring other dosimetric methods for simpler, quicker NAD analysis. Procedural recommendations include better documentation of conversion factor (activity-to-fluence and fluence-to-dose) derivations and spectrum uses, and updated analysis spreadsheets or simple Graphic User Interfaces for dose calculations. In conclusion, programmatic recommendations include formalized training for NAD analysts, and having multiple SMEs trained on the NAD program.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS

Heat transfer coefficients of moving particle beds from flow-dependent thermal conductivity and near-wall resistance

Accurate determination of heat transfer coefficients for flowing packed particle beds is essential to the design of particle heat exchangers and other thermal and thermochemical equipment. While such dense granular flows mostly fall into the well-known plug-flow regime, the discrete nature of granular materials alters the thermal transport processes in both the near-wall and bulk regions of flowing particle beds from their stationary counterparts. As a result, heat transfer correlations based on the stationary particle bed thermal conductivity could be inadequate for flowing particles in a heat exchanger. Most earlier works have achieved a reasonable agreement with experiments by treating granular heat transfer media as a plug-flow continuum with a near-wall thermal resistance in series. However, the thermal conductivity values of the continuum were often obtained from measurements on stationary beds owing to the difficulty of flowing bed measurements. In this work, it was found that the properties of a stationary bed are highly sensitive to the method of particle packing and there is a decrease in the particle bed thermal conductivity and increase in the near-wall thermal resistance, measured as an effective air gap thickness, on the onset of particle flow. These variations in thermal conductivity of stationary and flowing particle beds can lead to errors in heat transfer coefficient calculations. Therefore, the heat transfer coefficients for granular flows were calculated using experimentally determined flowing particle bed thermal conductivity and near-wall air gap for ceramic particles – CARBO CP 40/100 (mean diameter = 275 µm), HSP 40/70 (404 µm) and HSP 16/30 (956 µm); at velocities of 5–15 mm·s –1 ; and temperatures of 300–650 °C. The thermal conductivity and air gap values for CP 40/100 and HSP 40/70 were further used to calculate heat transfer coefficients across different particle bed temperatures and velocities for different parallel-plate heat exchanger dimensions. Furthermore, these calculations, which show good agreement with measured HTC values reported in literature, can be used as a guide for heat exchanger designs. Graphical abstract

14 SOLAR ENERGY

GX: a GPU-native gyrokinetic turbulence code for tokamak and stellarator design

GX is a code designed to solve the nonlinear gyrokinetic system for low-frequency turbulence in magnetized plasmas, particularly tokamaks and stellarators. In GX, our primary motivation and target is a fast gyrokinetic solver that can be used for fusion reactor design and optimization along with wide-ranging physics exploration. Here, this has led to several code and algorithm design decisions, specifically chosen to prioritize time to solution. First, we have used a discretization algorithm that is pseudospectral in the entire phase space, including a Laguerre–Hermite pseudospectral formulation of velocity space, which allows for smooth interpolation between coarse gyrofluid-like resolutions and finer conventional gyrokinetic resolutions and efficient evaluation of a model collision operator. Additionally, we have built GX to natively target graphics processors (GPUs), which are among the fastest computational platforms available today. Finally, we have taken advantage of the reactor-relevant limit of small $\rho _*$ by using the radially local flux-tube approach. In this paper we present details about the gyrokinetic system and the numerical algorithms used in GX to solve the system. We then present several numerical benchmarks against established gyrokinetic codes in both tokamak and stellarator magnetic geometries to verify that GX correctly simulates gyrokinetic turbulence in the small $\rho _*$. Moreover, we show that the convergence properties of the Laguerre–Hermite spectral velocity formulation are quite favourable for nonlinear problems of interest. Coupled with GPU acceleration, which we also investigate with scaling studies, this enables GX to be able to produce useful turbulence simulations in minutes on one (or a few) GPUs and higher fidelity results in a few hours using several GPUs. GX is open-source software that is ready for fusion reactor design studies.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY

CHARMM-GUI Bicelle Builder : An Extension of Membrane Builder for Modeling and Simulation of Bicelle Systems

Membrane mimetics, such as detergent micelles, nanodiscs, and amphipol complexes, which can provide membrane-like environments while retaining small and soluble features, have been utilized to study membrane proteins. A bicelle, composed of varying lipids and detergents, is a useful membrane mimetic because the lipid-to-detergent ratio, the q-value, can be adjusted to alter the properties of the aggregate, including the thickness and size of the bicelle. However, building a bicelle model for modeling and simulation studies requires nontrivial efforts, even for experts. We introduce CHARMM-GUI Bicelle Builder, a web-based platform that can generate various all-atom bicelle systems via a graphical user interface with all available lipids and detergents in Membrane Builder. To illustrate and validate Bicelle Builder with practical systems, we have modeled and simulated pure bicelles consisting of 1,2-dimyristoyl-sn-glycero-3-phosphocholine (DMPC) lipids with 1,2-dihexanoyl-sn-glycero-3-phosphocholine (C6DHPC) detergents and protein–bicelle complexes, composed of DMPC with C6DHPC, foscholine-10 (FOS10), and lysophosphatidylcholine-12 (LPC12) detergents. Our simulation results indicate that Bicelle Builder can generate reliable and robust bicelle models with and without proteins that retain DMPC bilayer characteristics. Bicelle Builder is expected to help researchers better understand not only bicelles themselves but also atomistic-level structures of protein–bicelle complexes that are often difficult to access through experimental approaches.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Nanopolysaccharide Builder: A User-Friendly Tool for Atomistic Models of Polysaccharide-Based Nanostructures

Here, we introduce Nanopolysaccharide Builder (NPB), a user-friendly software tool designed to construct polysaccharide nanostructures─mainly those based on cellulose, chitin, and chitosan─using experimental data or user-defined parameters. NPB enables the generation of cellulose and chitin allomorphs with customizable biochemical topologies and also facilitates the construction of large bundles that replicate nanostructures found in biological support systems, including plant cell walls and arthropod cuticles. The software outputs atomic Cartesian coordinates in Protein Data Bank (PDB) format and also provides atom connectivity files in PSF and PARM formats, ensuring seamless integration with major molecular dynamics (MD) engines such as NAMD, CHARMM, GROMACS, AMBER, OpenMM, and LAMMPS. Built on an interactive visualization framework, NPB features a graphical user interface (GUI) and supports both macOS and Linux operating systems. By enabling detailed atomic-scale studies of polysaccharide evolution in extracellular matrices and cell walls of algae, bacteria, fungi, and plants, NPB is poised to advance AI-guided research in sustainable chemical development and biomass utilization.

Wan, Zhangmin [Univ. of British Columbia, Vancouve