Search NASA⌕ Search

SEARCH · Search NASA

Results for “Computing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 739 records · Page 41

Toward a persistent event-streaming system for high-performance computing applications

High-performance computing (HPC) applications have traditionally relied on parallel file systems and file transfer services to manage data movement and storage. Alternative approaches have been proposed that use direct communications between application components, trading persistence and fault tolerance for speed. Event-driven architectures, as popularized in enterprise contexts, present a compelling middle ground, avoiding the performance cost and API constraints of parallel file systems while retaining persistence and offering impedance matching between application components. However, adapting streaming frameworks to HPC workloads requires addressing challenges unique to HPC systems. This paper investigates the potential for a streaming framework designed for HPC infrastructures and use cases. We introduce Mofka, a persistent event-streaming framework designed specifically for HPC environments. Mofka combines the capabilities of a traditional streaming service with optimizations tailored to the HPC context, such as support for massively multicore nodes, efficient scaling for large producer-consumer workflows, RDMA-enabled high-performance network communications, specialized network fabrics with multiple links per node, and efficient handling of large scientific data payloads. Built using the Mochi suite of HPC data service components, Mofka provides a lightweight, modular, and high-performance solution for persistent streaming in HPC systems. We present the architecture of Mofka and evaluate its performance against Kafka and Redpanda using benchmarks on diverse platforms, including Argonne's Polaris and Oak Ridge's Frontier supercomputers, showing up to 8× improvement in throughput in some scenarios. We then demonstrate its utility in several real-world applications: a tomographic reconstruction pipeline, a workflow for the discovery of metal-organic frameworks for carbon capture, and the instrumentation of Dask workflows for provenance tracking and performance analysis.

HPC↗

Experimental and Computational Characterization of a Modified Sioutas Cascade Impactor for Respirable Radioactive Aerosols

Oak Ridge National Laboratory is collecting and characterizing aerosols released when spent nuclear fuel (SNF) rods are fractured in bending. An aerosol collection system was designed and tested to collect respirable sized (<10 μm aerodynamic diameter [AED]) particulates inside a hot cell facility. The setup is a modified version of the commercially available Sioutas cascade impactor, to which additional stages were added to expand the aerosol collection range from 2.5 to ~15 μm AED. To accommodate the additional stages and specific test conditions, the operating flow rate for aerosol collection was reduced, and testing was conducted by using pressure drop measurements, surrogate dust collection, and particle size characterization. The fluid flow distribution within the cascade and its stages was simulated in STAR-CCM+, and the stage-wise pressure drops obtained using the computational fluid dynamics model were then compared to experimental data. Lagrangian particle simulations were also performed, and stage-wise collection statistics were obtained from the simulation for comparison with the experimental data obtained using SNF-surrogate dust particles. The results provide valuable insights into the stage-wise particle collection characteristics of the modified cascade impactor and can also be used to improve the prediction accuracy of the manufacturer-determined analytical correlations.

aerosol modeling↗

In Search of Optimum Fresh-Cut Raw Material: Using Computer Vision Systems as a Sensory Screening Tool for Browning-Resistant Romaine Lettuce Accessions

The popularity of ready-to-eat (RTE) salads has prompted novel technology to prolong the shelf life of their ingredients. Fresh-cut romaine lettuce is widely used in RTE salads; however, its tendency to quickly discolor continues to be a challenge for the industry. Selecting the ideal lettuce accessions for use in RTE salads is essential to ensure maximum shelf life, and it is critical to have a practical way to assess and compare the quality of multiple lettuce accessions that are being considered for use in fresh-cut applications. Thus, in this work we aimed to determine whether a computer vision system (CVS) composed of image acquisition, processing, and analysis could be effective to detect visual quality differences among 16 accessions of fresh-cut romaine lettuce during postharvest storage. The CVS involved a post-capturing color correction, effective image segmentation, and calculation of a browning index, which was tested as a predictor of quality and shelf life of fresh-cut romaine lettuce. The results demonstrated that machine vision software can be implemented to replace or supplement the scoring of a trained panel and instrumental quality measurements. Overall visual quality, a key sensory parameter that determines food preferences and consumer behavior, was highly correlated with the browning index, with a Pearson correlation coefficient of −0.85. Other important sensory decision parameters were also strongly or moderately correlated with the browning index, with Pearson correlation coefficients of −0.84 for freshness, 0.79 for off odor, and 0.57 for browning. The ranking of the accessions according to quality acceptability from the sensory evaluation produced a similar pattern to those obtained with the CVS. This study revealed that multiple lettuce accessions can be effectively benchmarked for their performance as fresh-cut sources via a CVS-based method. Future opportunities and challenges in using machine vision image processing to predict consumer preferences for RTE salad greens is also discussed.

Agriculture↗

Computationally efficient method for determining limiting velocities of edge dislocations in anisotropic crystals

The continuum-limit theory of dislocations in crystals predicts divergences in the elastic energy at crystal-geometry dependent limiting velocities vL, which separate subsonic, transsonic, and supersonic dislocation glide regimes and are therefore import for material strength models at high strain rates. Although it is known how to calculate those limiting velocities, there is one special case - edge dislocations with reflection symmetry, but non-vanishing elastic constants c16 or c26 - where previous methods have been notoriously slow. In this letter, we address this deficiency by deriving a computationally efficient method for determining the limiting velocities of edge dislocations with reflection symmetry which is two orders of magnitude faster than the previous method.

36 MATERIALS SCIENCE↗

The Melting Behavior of Hydrogen Direct Reduced Iron in Molten Steel and Slag: An Integrated Computational and Experimental Study

Direct reduced iron (DRI) and hot briquetted iron (HBI) are essential feedstocks for tramp element control in the electric arc furnace (EAF). Due to greenhouse gas (GHG) concerns related to CO2 emissions, hydrogen as a substitute for natural gas and a reductant in DRI production is being widely explored to reduce GHG emissions in ironmaking. This study examines the melting behavior of hydrogen DRI (H-DRI) pellets in the EAF containing low-carbon (0.1 wt.%) molten steel and molten slag. A computational heat transfer model was developed to predict the melting behavior of H-DRI pellets. To validate the model, a set of experimental laboratory simulations was conducted by immersing H-DRI in a molten steel bath and slag. The temperature history at the center of the pellet during melting and the shell thickness at different melting stages were utilized to validate the model. The simulation results agree with the experimental measurements of steel balls and H-DRI in different metallic molten steel and slag baths.

Materials Science↗

In-situ mid-circuit qubit measurement and reset in a single-species trapped-ion quantum computing system

We implement in-situ mid-circuit measurement and reset (MCMR) operations on a trapped-ion quantum computing system by using metastable qubit states in $^{171}\textrm{Yb}^+$ ions. We introduce and compare two methods for isolating data qubits from measured qubits: one shelves the data qubit into the metastable state and the other drives the measured qubit to the metastable state without disturbing the other qubits. We experimentally demonstrate both methods on a crystal of two $^{171}\textrm{Yb}^+$ ions using both the $S_{1/2}$ ground state hyperfine clock qubit and the $S_{1/2}$-$D_{3/2}$ optical qubit. These MCMR methods result in errors on the data qubit of about $2\%$ without degrading the measurement fidelity. With straightforward reductions in laser noise, these errors can be suppressed to less than $0.1\%$. The demonstrated method allows MCMR to be performed in a single-species ion chain without shuttling or additional qubit-addressing optics, greatly simplifying the architecture.

Atomic Physics (physics.atom-ph)↗

Decentralized Distributed Proximal Policy Optimization (DD-PPO) for High Performance Computing Scheduling on Multi-User Systems

Resource allocation in High Performance Computing (HPC) environments presents a complex and multifaceted challenge for job scheduling algorithms. Beyond the efficient allocation of system resources, schedulers must account for and optimize multiple performance metrics, including job wait time and system throughput. Traditional heuristic-based scheduling algorithms increasingly struggle and lack the efficiency needed to meet the demands and address the complexity and scale of modern HPC systems. Consequently, recent research efforts have focused on leveraging advancements in Artificial Intelligence (AI) and Deep Learning (DL), particularly Reinforcement Learning (RL), to develop more adaptable and intelligent scheduling strategies. Previous RL-based scheduling approaches have explored a range of algorithms, from Deep Q-Networks (DQN) to Proximal Policy Optimization (PPO), and more recently, hybrid methods that integrate Graph Neural Networks (GNNs) with RL techniques. However, a common limitation across these methods is their reliance on relatively small datasets, with few methods being evaluated using large-scale, multi-million-job trace datasets representative of real-world HPC workloads. Moreover, existing RL schedulers face scalability issues due to centralized policy updates, which hinder training efficiency and performance when applied to large datasets. This study introduces a novel RL-based scheduler utilizing Decentralized Distributed Proximal Policy Optimization (DD-PPO) algorithm, which supports large-scale distributed training across multiple workers without requiring parameter synchronization at every step. By eliminating reliance on centralized updates to a shared policy, the DD-PPO scheduler enhances scalability, training efficiency, and sample utilization. Experimental validation using a large real-world dataset containing over 11.5 million job traces collected from petascale HPC systems over six years assesses the influence of dataset scale on training effectiveness and compares DD-PPO performance to traditional and advanced scheduling approaches. The experimental results demonstrate improved scheduling performance in comparison to both heuristic-based schedulers and existing RL-based scheduling algorithms.

AI↗

Efficiently Computable Limits on EPR Pair Generation in Quantum Broadcast Channels

We investigate the generation of EPR pairs between three observers in a general causally structured setting, where communication occurs via a noisy quantum broadcast channel. The most general quantum codes for this setup take the form of tripartite quantum channels. Since the receivers are constrained by causal ordering, additional temporal relationships naturally emerge between the parties. These causal constraints enforce intrinsic no-signalling conditions on any tripartite operation, ensuring that it constitutes a physically realizable quantum code for a quantum broadcast channel. We analyze these constraints and, more broadly, characterize the most general quantum codes for communication over such channels. We examine the capabilities of codes that are fully no-signalling among the three parties, positive partial transpose (PPT)-preserving, or both, and derive simple semidefinite programs to compute the achievable entanglement fidelity. We then establish a hierarchy of semidefinite programming converse bounds -- both weak and strong -- for the capacity of quantum broadcast channels for EPR pair generation, in both one-shot and asymptotic regimes. Notably, in the special case of a point-to-point channel, our strong converse bound recovers and strengthens existing results. Finally, we demonstrate how the PPT-preserving codes we develop can be leveraged to construct PPT-preserving entanglement combing schemes, and vice versa.

FOS: Physical sciences↗

Self-driving thin film laboratory: autonomous epitaxial atomic-layer synthesis via real-time computer vision analysis of electron diffraction

Emerging materials science platforms with the ability to make autonomous decisions on the fly are fundamentally changing the outlook and protocols for materials optimization and discovery. Because AI-driven self-navigating schemes can effectively reduce the total number of iterations needed to arrive at the "answer" (i.e. the best stochiometric composition for a desired physical property, optimum materials processing parameters, etc.) by significant margins, they have the potential to revolutionize materials and chemical manufacturing processes at large in research laboratory settings as well as in industrial plants. Here, we demonstrate a successful implementation of real-time closed-loop autonomous navigation of a multi-dimensional materials synthesis parameter space for fabricating phase-pure epitaxial films of a metastable phase of a functional oxide in a combinatorial pulsed laser deposition chamber. Sequential epitaxial growth iterations in search of the optimized recipe to stabilize the desired crystal phase were performed using frame-by-frame quantitative computer vision analysis of reflection high-energy electron diffraction (RHEED) images of the unit-cell level film being deposited. The autonomous scheme regularly resulted in > 30-fold reduction in the number of required experiments compared to a comprehensive mapping of the parameter space. The real-time workflow developed here can be readily extended to a variety of thin film synthesis platforms opening the door for self-driving atomic-level materials design as well as autonomous optimization of semiconductor manufacturing.

36 MATERIALS SCIENCE↗

Supporting Data: Synthesis and Computational Analysis of Uranium(III)-Pnictogen Bonds

This repository contains the supporting data for the publication titled "Synthesis and Computational Analysis of Uranium(III)-Pnictogen Bonds" by Lauren M. Lopez, Diana Perales, Allison N. Smolek, Matthias Zeller, Bess Vlaisavljevich and Suzanne C. Bart. The manuscript associated with these data was published in J. Am. Chem. Soc. (DOI: 10.1021/jacs.5c19802). For readers only seeking the XYZ files associated with the publication, a second zipped folder called "coordinates.tgz" is included. The README.txt file describes the contents of the folder containing the input and output files.

Vlaisavljevich, Bess [Department of Chemistry; Uni↗

Transforming the Bootstrap: Using Transformers to Compute Scattering Amplitudes in Planar N = 4 Super Yang-Mills Theory

We pursue the use of deep learning methods to improve state-of-the-art computations in theoretical high-energy physics. Planar N = 4 Super Yang-Mills theory is a close cousin to the theory that describes Higgs boson production at the Large Hadron Collider; its scattering amplitudes are large mathematical expressions containing integer coefficients. In this paper, we apply Transformers to predict these coefficients. The problem can be formulated in a language-like representation amenable to standard cross-entropy training objectives. We design two related experiments and show that the model achieves high accuracy (> 98%) on both tasks. Our work shows that Transformers can be applied successfully to problems in theoretical physics that require exact solutions.

Lance, Dixon↗

Distributed Multi-GPU Community Detection on Exascale Computing Platforms

Community detection is a fundamental operation in graph mining, and by uncovering hidden structures and patterns within complex systems it helps solve fundamental problems pertaining to social networks, such as information diffusion, epidemics, and recommender systems. Scaling graph algorithms for massive networks becomes challenging on modern distributed-memory multi-GPU (Graphics Processing Unit) systems due to limitations such as irregular memory access patterns, load imbalances, higher communication-computation ratios, and cross-platform support. We present a novel algorithm HiPDPL-GPU (Distributed Parallel Louvain) to address these challenges. We conduct experiments involving different partitioning techniques to achieve an optimized performance of HiPDPL-GPU on the two largest supercomputers: Frontier and Summit. Remarkably, HiPDPL-GPU processes a graph with 4.2 billion edges in less than 3 minutes using 1024 GPUs. Qualitatively, the performance of HiPDPL-GPU is similar or better compared to other state-of-the-art CPU- and GPU-based implementations. While prior GPU implementations have predominantly employed CUDA, our first-of-its-kind implementation for community detection is cross-platform, accommodating both AMD and NVIDIA GPUs.

Sattar, Naw Safrin↗

Non-Electricity Based Renewable Fuels: Theory and Computation for Solar Thermochemical Hydrogen

Dominated by photovoltaics and wind, current renewable energy sources generate mostly electricity, but 80% of the global final energy consumption occurs in form of fuels. Therefore, direct solar fuel generation would be a major breakthrough for the energy transition. Solar thermochemical hydrogen (STCH) is one of the very few potential routes towards scalable renewable fuels, but currently suffers from lack of an oxide working material that could optimally perform energy conversion within the thermodynamic boundary conditions. Theory and computation can contribute in two distinct ways, through materials search and discovery, but also by providing detailed mechanistic models for specific systems so to advance our understanding of possible design strategies. To enable high-throughput materials screening, we developed a defect graph neural network (dGNN) machine learning approach,[1] which accelerates the prediction of defect formation energies by replacing the tedious density functional theory (DFT) supercell calculations for all possible defect sites. This approach enables high-throughput database screening of oxides, which was integrated with thermodynamic modeling to extract the reduction entropies as additional selection criterion for STCH. Once potential candidate materials are identified, detailed models can guide materials design by predicting performance characteristics. One challenge is to quantitatively predict thermochemical equilibria at high concentrations when the redox active defects start to interact with each other, thereby impeding the formation of additional defects. Introducing a model for the free energy of defect interaction, parametrized on the basis of DFT data, we simulated the complete STCH redox cycle for (Sr,Ce)MnO3 alloys, achieving near-quantitative agreement with experimental data.[2] The analysis of these simulations reveals how defect interactions diminish the reduction entropy and H2 yield, suggesting to include these interactions in design considerations. Finally, we revisit the popular van't Hoff method for analyzing reduction enthalpies and entropies. This method is not ideal, as it involves a temperature-dependent convolution of gas-phase and solid-state entropies, causing uncertainties in the same order of magnitude as the physical quantities of interest. To avoid this problem, we suggest a simple alternative approach which can be applied to experimental and simulated data alike.

first-principles calculations↗

System, method, and computer program for creating geometry-compliant lattice structures

A system and method of creating a shape-conforming lattice structure for a part formed via additive manufacturing. The method includes receiving a computer model of the part and generating a finite element mesh. A lattice structure including a number of lattice cellular components may also be generated. Some of the mesh elements of the finite element mesh may be deformed so that the finite element mesh conforms to the overall shape of the part. The lattice structure may then be deformed so that the lattice structure has a cellular periodicity corresponding to the finite elements of the finite element mesh. In this way, the part retains the benefits of its overall shape and the benefits of lattice features without introducing structural weak points, directional stresses, and other structural deficiencies.

Vernon, Gregory John↗

Quantum computing structures and resonators thereof

Embodiments disclosed herein include a resonator for use in quantum computing. The resonator can include a housing that is disposed along a resonator axis. The housing can have a first portion extending from a housing distal end to near a qubit location and a second portion extending from near the qubit location to a housing proximal end. The housing can define a cavity extending from a cavity proximal end to a cavity distal end along a portion of the resonator axis. The housing can include a protrusion extending axially from the housing distal end along the resonator axis to near the qubit location. A proximal portion of the protrusion can include a tapered portion. The resonator can include a qubit extending into the cavity at the qubit location.

Kutsaev, Sergey↗

Cosmic Reionization on Computers: Statistical Properties of the Distributions of Mean Opacities

Quasar absorption lines provide a unique window to the relationship between galaxies and the intergalactic medium during the Epoch of Reionization. In particular, high redshift quasars enable measurements of the neutral hydrogen content of the universe. However, the limited sample size of observed quasar spectra, particularly at the highest redshifts, hampers our ability to fully characterize the intergalactic medium during this epoch from observations alone. In this work, we characterize the distributions of mean opacities of the intergalactic medium in simulations from the Cosmic Reionization on Computers (CROC) project. We find that the distribution of mean opacities along sightlines follows a non-trivial distribution that cannot be easily approximated by a known distribution. When comparing the cumulative distribution function of mean opacities measurements in subsamples of sample sizes similar to observational measurements from the literature, we find consistency between CROC and observations at redshifts $z\lesssim 5.7$. However, at higher redshifts ($z\gtrsim5.7$), the cumulative distribution function of mean opacities from CROC is notably narrower than those from observed quasar sightlines implying that observations probe a systematically more opaque intergalactic medium at higher redshifts than the intergalactic medium in CROC boxes at these same redshifts. This is consistent with previous analyses that indicate that the universe is reionized too early in CROC simulations.

79 ASTRONOMY AND ASTROPHYSICS↗

Computational Fluid Dynamics Simulations of Glass Vitrification Refractory Coupon Tests

The Waste Treatment and Immobilization Plant (WTP) at the Hanford site is nearing the start of the Direct-Feed Low-Activity Waste (DFLAW) operations. DFLAW is destined to convert a pretreated low activity waste portion of the 56 million gallons of tank waste into a stable solid glass. In the subsequent decade completion of the high-level waste (HLW) facility is anticipated. Sustained operational missions of both LAW and HLW melter facilities are expected over multiple decades. In high-temperature glass melters, the refractory lining corrodes over time, which could potentially be an issue for longer term operations, this refractory corrosion is higher at the level of the glass-air interface due to surface tension driven flow. The glass viscosity, melt pool temperature, and glass chemical composition can impact the rate at which the refractory corrodes. This rate is important to quantify for the various waste glasses to be produced at the WTP since the integrity of the refractory should not be a limiting factor affecting the lifetime of the melter. To this end, a series of glasses representative of the first batches of waste glass produced by the WTP will be melted in small-scale crucibles with Monofrax® K-3 coupons inserted. The corrosion of the K-3 will be measured in the melt and at the meltline (or neckline). A model for the corrosion rate will be constructed and implemented into a previously developed framework for a computational fluid dynamics (CFD) model of the full-scale WTP. To assist with experimental design and validate the implementation of the model in the full-scale melter, CFD simulations of the small-scale crucible tests were performed. The bubbling that occurs in the small-scale crucible is initially validated here with a model that uses silicone oil at room temperature. The viscosity of the oil ranges from 1 to 100 Pa•s, which corresponds to operating glass pool temperatures near 1150 °C down to idling temperatures near 950 °C. The simulation results show good agreement with the bubble sizes that form during experiments. CFD modeling of the crucible setup was used to determine bubbling characteristics to match the range of near-wall velocities expected in the full-scale WTP. This study presents the initial CFD modeling results, corrosion testing plan, and some preliminary corrosion samples with an outline for the next steps for the development of the corrosion model.

Abboud, Alexander W. [Idaho National Lab]↗