Search NASASearch

SEARCH · Search NASA

Results for “Edge computing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 199 records · Page 11

autoGEMM: Pushing the Limits of Irregular Matrix Multiplication on Arm Architectures

This paper presents an open-source library that pushes the limits of performance portability for irregular General Matrix Multiplication (GEMM) on the widely-used Arm architectures. Our library, autoGEMM, is designed to support a wide range of Arm processors: from edge devices to HPC-grade CPUs. autoGEMM generates optimized kernels for various hardware configurations by auto-combining fragments of autogenerated micro-kernels that employ hand-written optimizations to maximize computational efficiency. We optimize the kernel pipeline by tuning the register reuse and the data load/store overlapping. In addition, we use a dynamic tiling scheme to generate balanced tile shapes. Finally, we position autoGEMM on top of the TVM framework where our dynamic tiling scheme prunes the search space for TVM to identify the optimal combination of parameters for code optimization. Evaluations on five different classes of Arm chips demonstrate the advantages of autoGEMM. For small matrices, autoGEMM achieves 98% of peak and up to 2.0x speedup over state-of-the-art libraries such as LIBXSMM and LibShalom. For irregular matrices (i.e. tall skinny and long rectangles), autoGEMM is 1.3-2.0x faster than widely-used libraries such as OpenBLAS and Eigen. autoGEMM is publicly available at: https://github.com/wudu98/autoGEMM.

Wu, Du

Analysis of Second Target Station Target Removal Dose Rates

The Second Target Station (STS) project at Oak Ridge National Laboratory’s spallation neutron source is a crucial initiative for maintaining U.S. leadership in neutron sciences. The STS aims to create the world’s brightest pulsed cold neutron source, enabling cutting-edge research across various scientific disciplines. To ensure safe and efficient maintenance operations, understanding the effects of shutdown dose rates from activated components within the STS target systems is essential. This study establishes a computational framework for calculating decay gamma sources and subsequent shutdown dose rates utilizing advanced methods to account for all activation channels, including high-energy interactions down to thermal neutron capture. This study describes a novel integration of multiple tools and provides an effective means of analyzing activation and shutdown dose rates at spallation neutron facilities. A custom-developed script automates the decay gamma source generation process, ensuring proper sampling during the variance reduction phase, which is critical for accurate predictions of shutdown dose rates.

Transmutation

Exploring the relation between transonic dislocation glide and stacking fault width in FCC metals

Theory predicts limiting gliding velocities that dislocations cannot overcome. Computational and recent experiments have shown that these limiting velocities are soft barriers and dislocations can reach transonic speeds in high rate plastic deformation scenarios. In this paper we systematically examine the mobility of edge and screw dislocations in several face centered cubic (FCC) metals (Al, Au, Pt, and Ni) in the extreme large-applied-stress regime using molecular dynamics simulations. Our results show that edge dislocations are more likely to move at transonic velocities due to their high mobility and lower limiting velocity than screw dislocations. Importantly, among the considered FCC metals, the dislocation core structure determines the dislocation’s ability to reach transonic velocities. This is likely due to the variation in stacking fault width due to relativistic effects near the limiting velocities.

36 MATERIALS SCIENCE

Geometric representations of braid and Yang–Baxter gates

Brick-wall circuits composed of the Yang–Baxter gates are integrable. It becomes an important tool to study the quantum many-body system out of equilibrium. To put the Yang–Baxter gate on quantum computers, it has to be decomposed into the native gates of quantum computers. It is favorable to apply the least number of native two-qubit gates to construct the Yang–Baxter gate. We study the geometric representations of all X-type braid gates and their corresponding Yang–Baxter gates via the Yang–Baxterization. We find that the braid and Yang–Baxter gates can only exist on certain edges and faces of the two-qubit tetrahedron. We identify the parameters by which the braid and Yang–Baxter gates are the Clifford gate, the matchgate, and the dual-unitary gate. The geometric representations provide the optimal decompositions of the braid and Yang–Baxter gates in terms of other two-qubit gates. We also find that the entangling powers of the Yang–Baxter gates are determined by the spectral parameters. Our results provide the necessary conditions to construct the braid and Yang–Baxter gates on quantum computers.

97 MATHEMATICS AND COMPUTING

Towards fully predictive gyrokinetic full- f simulations: validation and triangularity studies in TCV

Designing economical magnetic confinement fusion power plants motivates computational tools that can estimate plasma behavior from engineering parameters without direct reliance on experimental measurement of the plasma profiles. In this work, we present full-f global long-wavelength gyrokinetic simulations of edge and scrape-off layer turbulence in tokamaks that use only magnetic geometry, heating power, and particle inventory as inputs. Unlike many modeling approaches that employ free parameters fitted to experimental data, raising uncertainties when extrapolating to reactor scales. This approach directly simulates turbulence and resulting profiles through gyrokinetics without such empirical adjustments. This is achieved via an adaptive sourcing algorithm in Gkeyll that strictly controls energy injection and emulates particle sourcing due to neutral recycling. We show that the simulated kinetic profiles compare reasonably well with Thomson scattering and Langmuir probe data for Tokamak á Configuration Variable (TCV) discharge #65125, and that the simulations reproduce characteristic features such as blob transport and self-organized electric fields. Applying the same framework to study triangularity effects suggests mechanisms contributing to the improved confinement reported for negative triangularity (NT). Simulations of TCV discharges #65125 and #65130 indicate that NT increases the E x B flow shear (by about 20% in these cases), which correlates with reduced turbulent losses and a modest change in the distribution of power exhaust to the vessel wall. While the physical models contain approximations that can be refined in future work, the predictive capability demonstrated here, evolving multiple profile relaxation times with kinetic electron and ion models in hundreds of GPU hours, indicates the feasibility of using Gkeyll to support design studies of fusion devices.

Hoffmann, Antoine Cyril David [Princeton Plasma Ph

Tunable Few-Layer van der Waals Crystals and Heterostructures as Emerging Energy and Quantum Materials (Final Technical Report)

2D and layered (van der Waals) semiconductors offer extraordinary opportunities for manipulating optically excited charge carriers, many-body excitations, and non-charge based quantum numbers. To date, research has focused on a limited group of materials, mostly transition metal dichalcogenides in the monolayer limit. Other van der Waals semiconductors, and especially few-layer to multilayer crystals and their heterostructures, carry large potential for the discovery of phenomena of interest for future energy and information technologies. But they remain largely unexplored, often due to a lack of access to high-quality materials and approaches for measuring their properties at the relevant scales. The goal of this project was to develop an EPSCoR-State/National Laboratory Partnership that addresses the challenges of preparing high-quality van der Waals semiconductors and of probing their structure, composition, and especially their optoelectronic and photonic properties, near the atomic scale using electron microscopy techniques. A central component of the project was the development of advanced methods for electron microscopy and electron-excited spectroscopy, taking advantage of unique samples as well as leading capabilities and expertise at the partner institutions. Efforts to advance leading-edge techniques was supported by ancillary developments, such as precision sample preparation for electron microscopy/spectroscopy, coordinated chemical imaging, and analytical electron microscopy. Experiments in materials synthesis and technique development were closely linked to theory and computation. The results obtained under this project yielded multifaceted benefits to the involved partners and their institutions, DOE-BES, the wider scientific community, and society at large, particularly in the State of Nebraska through dividends from knowledge and human capital generated under the project.

36 MATERIALS SCIENCE

Tungsten-dioxo single-site heterogeneous catalyst on carbon: synthesis, structure, and catalysis

This study investigates the application of a novel third-row metal, tungsten, to carbon-supported single-site metal-oxo heterogeneous catalysis. Tungsten is a green and earth-abundant metal, but an unexplored candidate in this role. The carbon (AC = activated carbon)-supported tungsten dioxo complex, AC/WO 2 was prepared via grafting of (DME)WO 2 Cl 2 (DME = 1,2-dimethoxyethane) onto high-surface-area activated carbon. AC/WO 2 was fully characterized by ICP-OES, XPS, EXAFS, XANES, SMART-EM, and DFT. W 4d 7/2 XPS and W L III -Edge XANES assign the oxidation state as W(VI), while EXAFS reveals two W=O double and two W–O single bonds at distances of 1.73 and 1.92 Å, respectively. These data align well with DFT computational results, supporting the structure as Carbon(–μ-O–) 2 M(=O) 2 . SMART-EM verifies that single W(VI) catalytic sites are bonded in an out-of-plane manner. The catalytic performance of air- and water-stable AC/WO 2 is compared to that of AC/MoO 2 . AC/WO 2 is more active and selective than the molybdenum analog in mediating alcohol dehydration of various substrates, and is recyclable. Notably, AC/WO 2 is an effective and recyclable catalyst for primary aliphatic alcohol dehydration and forms no dehydrogenation side products in contrast to AC/MoO 2 . However, AC/WO 2 is less effective in epoxidation and PET depolymerization. Overall, this work demonstrates the potential of carbon-supported third row metals for future studies.

02 PETROLEUM

Libration of hydroxyl groups in layered aluminum (oxy)hydroxides and other material analogs: insights from inelastic neutron scattering and theory

We analyzed the hydroxyl librational signatures of five structurally related aluminum (oxy)hydroxides, using inelastic neutron scattering (INS) and plane-wave lattice dynamics simulations. A clear trend across these aluminum-containing phases illustrates the relationship between hydrogen bonding, local atomic structure, and the spectral location and profile of the librational bands. The INS spectra have been compared to previous optical spectroscopy and computational studies, highlighting the complementary nature of the INS technique. Taking into account other structurally or chemically related material analogs, we have identified a correlation between a blueshift (to higher energy) of the upper librational band edge and the geometry of the hydrogen bond interactions, mirroring (with opposite correlation) the well-known redshift in the intramolecular O–H stretching energy with increasing hydrogen bond strength. For hydroxyl groups that do not participate in hydrogen bonding effectively, the bending librations occur at lower energies and hybridize with metal–oxygen lattice modes. Standard density functional theory approximations, including dispersion corrections, struggle to correctly predict vibrational frequencies of motions dominated by H but perform well for metal–oxygen modes, allowing us to make detailed mode assignments in several cases, including a demonstration of how layer-to-layer disorder in boehmite hydrogen bond orientations is reflected in the sharp but minor low energy peaks (at ∼70–80 meV) of the INS spectrum.

Wang, Hsiu-Wen [Oak Ridge National Laboratory (ORN

Nontrivial fusion of Majorana zero modes in interacting quantum-dot arrays

Motivated by recent experimental reports of Majorana zero modes (MZMs) in quantum-dot systems at the “sweet spot,” where the electronic hopping t ℎ is equal to the superconducting coupling Δ, we study the time-dependent spectroscopy corresponding to the nontrivial fusion of MZMs. The term “nontrivial” refers to the fusion of Majoranas from different original pairs of MZMs, each with well-defined parities. We employ an experimentally accessible time-dependent real-space local density-of-states (LDOS) method to investigate the nontrivial MZM fusion outcomes in canonical chains and in a Y-shaped array of interacting electrons. In the case of quantum-dot chains where two pairs of MZMs are initially disconnected, after fusion we find equal-height peaks in the electron and hole components of the LDOS, signaling nontrivial fusion into both the vacuum I and fermion Ψ channels with equal weight. For π-junction quantum-dot chains, where the superconducting phase has opposite signs on the left and right portions of the chain, after the nontrivial fusion we observed the formation of an exotic two-site MZM near the center of the chain, coexisting with another single-site MZM. Furthermore, we also studied the fusion of three MZMs in the Y-shaped geometry. In this case, after the fusion we observed the novel formation of another exotic multisite MZM, with properties depending on the connection and geometry of the central region of the Y-shaped quantum-dot array.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC

Improved Statistics for F-theory Standard Models

Much of the analysis of F-theory-based Standard Models boils down to computing cohomologies of line bundles on matter curves. By varying parameters one can degenerate such matter curves to singular ones, typically with many nodes, where the computation is combinatorial and straightforward. The question remains to relate the (a priori possibly smaller) value on the original curve to the singular one. In this work, we introduce some elementary techniques (pruning trees and removing interior edges) for simplifying the resulting nodal curves to a small collection of terminal ones that can be handled directly. When applied to the QSMs, these techniques yield optimal results in the sense that obtaining more precise answers would require currently unavailable information about the QSM geometries. This provides us with an opportunity to enhance the statistical bounds established in earlier research regarding the absence of vector-like exotics on the quark-doublet curve.

Bies, Martin

Non-Covalent Interactions and Helical Packing in Thiophene-Phenylene Copolymers: Tuning Solid-State Ordering and Charge Transport for Organic Field-Effect Transistors

In this study, we introduce two thiophene-phenylene-thiophene (TPT) polymers designed to leverage noncovalent intramolecular interactions to regulate main-chain conformation and enhance solid-state ordering. By incorporating unsubstituted thiophene (T) or bithiophene (2T) units, we reveal striking divergence in the thermal, morphological, and optoelectronic properties of the resulting films, facilitated by these noncovalent interactions. Using a combination of computational and experimental approaches, we show that annealing yields remarkably different polymer conformations and, consequently, charge transport properties. TPT-T undergoes a significant structural transformation, adopting a more planar backbone conformation and a highly crystalline, edge-on molecular orientation. In contrast, the introduction of a single additional thiophene unit in TPT-2T leads to a more isotropic molecular orientation with a slight preference for face-on alignment, resulting in a heterogeneous film structure that hinders charge transport despite achieving tighter molecular packing. Remarkably, despite being composed of achiral components, TPT-2T develops chirality upon annealing, indicating the formation of a helical conformation. Organic field-effect transistor measurements reveal that the well-ordered alignment in annealed TPT-T films results in higher charge carrier mobility and a narrower distribution of mobility values than in TPT-2T. These findings provide critical insights into the structure−property relationships of conjugated polymers, offering guidance for optimizing molecular design and processing strategies for highperformance organic electronic materials.

36 MATERIALS SCIENCE

Nanometer Scale Imaging to Develop Quantitative Descriptors of Bipolar Membrane Junction Structure

Swings in pH can be achieved by electrically polarizing a bipolar membrane (BPM) to drive water dissociation at the BPM junction for electrochemical conversion and separation processes. BPM junction design is critical to tailor performance for specific applications; however, characterization techniques capable of resolving the nanometer scale physical structure of the junction are limited. We present sample preparation, imaging, and analysis workflows that are adaptable to a variety of BPM junction architectures. Atomic force microscopy produces BPM junction images with nanometer scale lateral resolution for samples with and without a graphene oxide water dissociation catalyst in the junction. Subsequent image segmentation and analysis quantify line edge roughness and catalyst layer thickness as descriptors of junction structure. Comparison of pre- and post-electrodialysis junctions suggests electric field-induced alignment of catalyst particles during electrodialysis. This characterization workflow can inform manufacturing protocols, computational modeling, and failure mode analysis for next-generation BPMs.

97 MATHEMATICS AND COMPUTING

DIMPLES: Distributed Influence Maximization for Pandemic pLanning on Exascale Systems

We study exascale parallel algorithms for the selection of intervention or monitoring strategies in massive realistic socio-technical networks through scalable Influence Maximization (InfMax) algorithms. We employ novel techniques to enable efficient scaling on up to 8k nodes of OLCF Frontier, with 65k AMD GPUs and 458k AMD CPU cores. Current state-of-the-art InfMax tools are limited to networks with only a few million actors (vertices) and a few hundred million interactions (edges). By overcoming these limitations, we show that our approach is capable of processing a realistic social contact network of the United States with 285 million nodes and about 8 billion edges. This two orders-of-magnitude improvement over the previous state-of-the-art is obtained by leveraging algorithmic advancements for the InfMax problem and designing several problem-specific approaches to overlap communication with computation, improve GPU efficiency, and lower the application’s memory requirements. We evaluate strong scaling for computing 10k most influential seeds using up to 8k nodes of an exascale system, and weak scaling from 128 to 8k system nodes for seed sets ranging from 625 to 40k seeds. We achieve the fastest-known runtime of 25 minutes while performing 48 million diffusion simulations totaling 2.31 petabytes to identify 40k influential seeds using 8k nodes, and take 5.75 minutes to identify 10k seeds while using 4k nodes.

Minutoli, Marco [Pacific Northwest National Labora

Picasso: Memory-Efficient Graph Coloring Using Palettes With Applications in Quantum Computing

A coloring of a graph is an assignment of colors to vertices such that no two neighboring vertices have the same color. The need for memory-efficient coloring algorithms is motivated by their application in computing clique partitions of graphs arising in quantum computations where the objective is to map a large set of Pauli strings into a compact set of unitaries. We present Picasso, a randomized memory-efficient iterative parallel graph coloring algorithm with theoretical sublinear space guarantees under practical assumptions. The parameters of our algorithm provide a trade-off between coloring quality and resource consumption. To assist the user, we also propose a machine learning model to predict the coloring algorithm’s parameters considering these trade-offs. We provide a sequential and a parallel implementation of the proposed algorithm. We perform an experimental evaluation on a 64-core AMD CPU equipped with 512 GB of memory and an Nvidia A100 GPU with 40GB of memory. For a small dataset where existing coloring algorithms can be executed within the 512 GB memory budget, we show up to 68× memory savings. On massive datasets we demonstrate that GPU-accelerated Picasso can process inputs with 49.5× more Pauli strings (vertex set in our graph) and 2,478× more edges than state-of-the-art parallel approaches.

artificial intelligence, quantum computing

Spectrally accelerated edge and scrape-off layer gyrokinetic turbulence simulations

This paper presents the first gyrokinetic (GK) simulations of edge and scrape-off layer (SOL) turbulence accelerated by a velocity-space spectral approach in the full-f GK code GENE-X. Building upon the original grid velocity-space discretization, we derive and implement a new spectral formulation and verify the numerical implementation using the method of manufactured solution. We conduct a series of spectral turbulence simulations focusing on the TCV-X21 reference case (Oliveira et al., 2022 [26]) and compare these results with previously validated grid simulations (Ulbl et al., 2023 [25]). The spectral approach reproduces the outboard midplane (OMP) profiles (density, temperature, and radial electric field), dominated by trapped electron mode (TEM) turbulence, with excellent agreement and significantly lower velocity-space resolution. As a consequence, the spectral approach reduces the computational cost (CPUh) by at least an order of magnitude, of approximately 50 for the TCV-X21 case. This enables high-fidelity GK simulations to be performed within a few days on modern CPU-based supercomputers for medium-sized devices and establishes GENE-X as a powerful tool for studying edge and SOL turbulence, moving towards reactor-relevant devices like ITER.

Gyrokinetic

Realization of fermionic Laughlin state on a quantum processor

Strongly correlated topological phases of matter are central to modern condensed matter physics and quantum information technology but often challenging to probe and control in material systems. The experimental difficulty of accessing these phases has motivated the use of engineered quantum platforms for simulation and manipulation of exotic topological states. Among these, the Laughlin state stands as a cornerstone for topological matter, embodying fractionalization, anyonic excitations, and incompressibility. Although its bosonic analogs have been realized on programmable quantum simulators, a genuine fermionic Laughlin state has yet to be demonstrated on a quantum processor. Here, we realize the ν = 1/3 fermionic Laughlin state on IonQ’s trapped-ion quantum computer using an efficient and scalable Hamiltonian variational ansatz with 369 two-qubit gates on a 16-qubit circuit. Employing symmetry-verification error mitigation, we extract key observables that characterize the Laughlin state, including correlation hole, bulk-edge correspondence, and topological entanglement entropy, with strong agreement to exact diagonalization benchmarks. This work demonstrates an end-to-end workflow to simulate material-intrinsic topological orders and provides a starting point to explore its dynamics and excitations on digital quantum processors.

Shen, Lingnan [Univ. of Washington, Seattle, WA (U

Drift-kinetic effects of tungsten on plasma response to RMP in ITER

Here, effects of high- Z ( Z is the particle charge number) tungsten impurity ions on the plasma response to the resonant magnetic perturbation (RMP) field are numerically investigated for the ITER 15 MA baseline scenario, where the tungsten contribution to the plasma response is computed with a drift-kinetic model while the bulk thermal particle contributions follow the fluid approximation. The study yields three highlights: (i) the drift-kinetic contribution of the tungsten impurity exerts minor influence on the plasma response compared to that computed by the pure fluid model without tungsten; (ii) a new figure of merit, based on the resonant spectrum perturbation at the plasma boundary surface, results in different optimal coil phasing compared to that previously obtained by maximizing the edge-peeling plasma response; (iii) the optimal $n = 3$ RMP (for edge localized mode (ELM) control, $n$ is the toroidal mode number) is found to induce a large tungsten particle influx near the plasma edge associated with the neoclassical toroidal viscosity. The study thus provides useful data on the compatibility of the full tungsten wall with RMP ELM control in ITER.

ITER