Algebraic Multigrid: A Meeting of Math and Computer Science
Talk at UNM 20th Annual Computer Science Student Conference 2025
SEARCH · Search NASA
Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.
Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.
Talk at UNM 20th Annual Computer Science Student Conference 2025
The Radiological Safety Analysis Computer (RSAC) Program Version 7.4 (RSAC-7) is the newest version of the RSAC legacy code. RSAC-7 calculates the consequences of a release of radionuclides to the atmosphere. Users generates a fission product inventory from either reactor operating history or a nuclear criticality event. RSAC-7 models the effects of high-efficiency particulate air filters or other cleanup systems and calculates the decay and ingrowth during transport through processes, facilities, and the environment. Doses are calculated for inhalation, air immersion, ground surface, ingestion, and cloud gamma pathways. RSAC-7 is used as a tool to evaluate accident conditions in emergency response scenarios, radiological sabotage events, and safety basis accident consequences. This users’ manual contains the mathematical models and operating instructions for RSAC-7. Instructions, screens, and examples are provided to guide the user through the functions provided by RSAC-7. This program is designed for users who are familiar with radiological dose assessment methods.
Presentation of my work this summer speeding up forward simulation of noisy quantum computers.
This talk will highlight recent efforts at the SQMS Center to develop qudit-based quantum computing architectures using superconducting three-dimensional (3D) cavities, as well as the use of these ultra-coherent cavities for quantum sensing. I will present systematic studies of materials and devices aimed at identifying and mitigating the dominant sources of decoherence—including two-level systems (TLS), quasiparticles, and other noise mechanisms—in both transmons and 3D cavities. These investigations include microwave loss characterization of niobium, tantalum, aluminum, their native oxides, and substrate materials such as silicon and sapphire. By combining measurements on qubits and cavities, we disentangle subsystem-specific loss mechanisms and establish a hierarchy of mitigation strategies, leading to transmon coherence times exceeding one millisecond. I will also discuss studies of quasiparticle dynamics, including quasiparticle bursts observed in qubits operated both above ground and at the Gran Sasso underground laboratory, and the observation that applied magnetic fields can suppress temporal T₁ fluctuations. Building on these advances, we demonstrate a record-coherence two-cell cavity-qudit system with coherence times exceeding 20 milliseconds. Leveraging tunable sideband interactions together with error-resilient protocols, including measurement-based error correction and post-selection, we achieve high-fidelity quantum state control, including the preparation of Fock states up to N=20 with fidelities above 95% and the generation of high-fidelity two-mode entangled states. Finally, I will discuss how these ultra-coherent quantum systems are enabling emerging quantum sensing applications, including searches for dark matter and gravitational waves.
I show how to compute the nonlinear power spectrum across the entire $w(z)$ dynamical dark energy model space. Using synthetic ΛCDM data, I train a neural ordinary differential equation (ODE) to infer the evolution of the nonlinear matter power spectrum as a function of the background expansion and mean matter density across ∼9 Gyr of cosmic evolution. After training, the model generalises to any dynamical dark energy model parameterised by $w(z)$. With little optimisation, the neural ODE is accurate to within 4% up to $k = 5\, h\, {\mathrm Mpc}^{−1}$. Unlike simulation rescaling methods, neural ODEs naturally extend to summary statistics beyond the power spectrum that are sensitive to the growth history.
High-performance computing (HPC) applications have traditionally relied on parallel file systems and file transfer services to manage data movement and storage. Alternative approaches have been proposed that use direct communications between application components, trading persistence and fault tolerance for speed. Event-driven architectures, as popularized in enterprise contexts, present a compelling middle ground, avoiding the performance cost and API constraints of parallel file systems while retaining persistence and offering impedance matching between application components. However, adapting streaming frameworks to HPC workloads requires addressing challenges unique to HPC systems. This paper investigates the potential for a streaming framework designed for HPC infrastructures and use cases. We introduce Mofka, a persistent event-streaming framework designed specifically for HPC environments. Mofka combines the capabilities of a traditional streaming service with optimizations tailored to the HPC context, such as support for massively multicore nodes, efficient scaling for large producer-consumer workflows, RDMA-enabled high-performance network communications, specialized network fabrics with multiple links per node, and efficient handling of large scientific data payloads. Built using the Mochi suite of HPC data service components, Mofka provides a lightweight, modular, and high-performance solution for persistent streaming in HPC systems. We present the architecture of Mofka and evaluate its performance against Kafka and Redpanda using benchmarks on diverse platforms, including Argonne's Polaris and Oak Ridge's Frontier supercomputers, showing up to 8× improvement in throughput in some scenarios. We then demonstrate its utility in several real-world applications: a tomographic reconstruction pipeline, a workflow for the discovery of metal-organic frameworks for carbon capture, and the instrumentation of Dask workflows for provenance tracking and performance analysis.
The occurrence of vaso-occlusive crisis greatly depends on the competition between the sickling delay time and the transit time of individual sickle cells, i.e., red blood cells from sickle cell disease (SCD) patients, while they are traversing the circulatory system. Many drugs for treating SCD work by inhibiting the polymerization of sickle hemoglobin (HbS), effectively delaying the sickling process in sickle cells (SS RBCs). Most previous studies on screening anti-sickling drugs, such as voxelotor, rely on in vitro testing of sickling characteristics, often conducted under prolonged deoxygenation for up to 1 hour. However, since the microcirculation of RBCs typically takes less than 1 minute, the results of these studies may be less accurate and less relevant for in vitro-in vivo correlation. In our current study, we introduce a computer vision-enhanced microfluidic framework designed to automatically capture the transient sickling kinetics of SS RBCs within a 1-min timeframe. Our study has successfully detected differences in the transient sickling kinetics between vehicle control and voxelotor-treated SS RBCs. This approach has the potential for broader applications in screening anti-sickling therapies.
Oak Ridge National Laboratory is collecting and characterizing aerosols released when spent nuclear fuel (SNF) rods are fractured in bending. An aerosol collection system was designed and tested to collect respirable sized (<10 μm aerodynamic diameter [AED]) particulates inside a hot cell facility. The setup is a modified version of the commercially available Sioutas cascade impactor, to which additional stages were added to expand the aerosol collection range from 2.5 to ~15 μm AED. To accommodate the additional stages and specific test conditions, the operating flow rate for aerosol collection was reduced, and testing was conducted by using pressure drop measurements, surrogate dust collection, and particle size characterization. The fluid flow distribution within the cascade and its stages was simulated in STAR-CCM+, and the stage-wise pressure drops obtained using the computational fluid dynamics model were then compared to experimental data. Lagrangian particle simulations were also performed, and stage-wise collection statistics were obtained from the simulation for comparison with the experimental data obtained using SNF-surrogate dust particles. The results provide valuable insights into the stage-wise particle collection characteristics of the modified cascade impactor and can also be used to improve the prediction accuracy of the manufacturer-determined analytical correlations.
The popularity of ready-to-eat (RTE) salads has prompted novel technology to prolong the shelf life of their ingredients. Fresh-cut romaine lettuce is widely used in RTE salads; however, its tendency to quickly discolor continues to be a challenge for the industry. Selecting the ideal lettuce accessions for use in RTE salads is essential to ensure maximum shelf life, and it is critical to have a practical way to assess and compare the quality of multiple lettuce accessions that are being considered for use in fresh-cut applications. Thus, in this work we aimed to determine whether a computer vision system (CVS) composed of image acquisition, processing, and analysis could be effective to detect visual quality differences among 16 accessions of fresh-cut romaine lettuce during postharvest storage. The CVS involved a post-capturing color correction, effective image segmentation, and calculation of a browning index, which was tested as a predictor of quality and shelf life of fresh-cut romaine lettuce. The results demonstrated that machine vision software can be implemented to replace or supplement the scoring of a trained panel and instrumental quality measurements. Overall visual quality, a key sensory parameter that determines food preferences and consumer behavior, was highly correlated with the browning index, with a Pearson correlation coefficient of −0.85. Other important sensory decision parameters were also strongly or moderately correlated with the browning index, with Pearson correlation coefficients of −0.84 for freshness, 0.79 for off odor, and 0.57 for browning. The ranking of the accessions according to quality acceptability from the sensory evaluation produced a similar pattern to those obtained with the CVS. This study revealed that multiple lettuce accessions can be effectively benchmarked for their performance as fresh-cut sources via a CVS-based method. Future opportunities and challenges in using machine vision image processing to predict consumer preferences for RTE salad greens is also discussed.
The continuum-limit theory of dislocations in crystals predicts divergences in the elastic energy at crystal-geometry dependent limiting velocities vL, which separate subsonic, transsonic, and supersonic dislocation glide regimes and are therefore import for material strength models at high strain rates. Although it is known how to calculate those limiting velocities, there is one special case - edge dislocations with reflection symmetry, but non-vanishing elastic constants c16 or c26 - where previous methods have been notoriously slow. In this letter, we address this deficiency by deriving a computationally efficient method for determining the limiting velocities of edge dislocations with reflection symmetry which is two orders of magnitude faster than the previous method.
Direct reduced iron (DRI) and hot briquetted iron (HBI) are essential feedstocks for tramp element control in the electric arc furnace (EAF). Due to greenhouse gas (GHG) concerns related to CO2 emissions, hydrogen as a substitute for natural gas and a reductant in DRI production is being widely explored to reduce GHG emissions in ironmaking. This study examines the melting behavior of hydrogen DRI (H-DRI) pellets in the EAF containing low-carbon (0.1 wt.%) molten steel and molten slag. A computational heat transfer model was developed to predict the melting behavior of H-DRI pellets. To validate the model, a set of experimental laboratory simulations was conducted by immersing H-DRI in a molten steel bath and slag. The temperature history at the center of the pellet during melting and the shell thickness at different melting stages were utilized to validate the model. The simulation results agree with the experimental measurements of steel balls and H-DRI in different metallic molten steel and slag baths.
Lyman limit systems (LLSs) are dense hydrogen clouds with high enough H i column densities to absorb Lyman continuum photons emitted from distant quasars. Their high column densities imply an origin in dense environments; however, the statistics and distribution of LLSs at high redshifts still remain uncertain. In this paper, we use self-consistent radiative transfer cosmological simulations from the Cosmic Reionization on Computers (CROC) project to study the physical properties of LLSs at the tail end of cosmic reionization at z ~ 6. We generate 3000 synthetic quasar sight lines to obtain a large number of LLS samples in the simulations. In addition, with the high physical fidelity and resolution of CROC, we are able to quantify the association between these LLS samples and nearby galaxies. Our results show that the fraction of LLSs spatially associated with nearby galaxies increases with H i column density. Moreover, we find that LLSs that are not near any galaxy typically reside in filamentary structures connecting neighboring galaxies in the intergalactic medium (IGM). This quantification of the distribution and association of LLSs to large-scale structure informs our understanding of the IGM–galaxy connection during the "Epoch of Reionization," and provides a theoretical basis for interpreting future observations.
We implement in-situ mid-circuit measurement and reset (MCMR) operations on a trapped-ion quantum computing system by using metastable qubit states in $^{171}\textrm{Yb}^+$ ions. We introduce and compare two methods for isolating data qubits from measured qubits: one shelves the data qubit into the metastable state and the other drives the measured qubit to the metastable state without disturbing the other qubits. We experimentally demonstrate both methods on a crystal of two $^{171}\textrm{Yb}^+$ ions using both the $S_{1/2}$ ground state hyperfine clock qubit and the $S_{1/2}$-$D_{3/2}$ optical qubit. These MCMR methods result in errors on the data qubit of about $2\%$ without degrading the measurement fidelity. With straightforward reductions in laser noise, these errors can be suppressed to less than $0.1\%$. The demonstrated method allows MCMR to be performed in a single-species ion chain without shuttling or additional qubit-addressing optics, greatly simplifying the architecture.
Resource allocation in High Performance Computing (HPC) environments presents a complex and multifaceted challenge for job scheduling algorithms. Beyond the efficient allocation of system resources, schedulers must account for and optimize multiple performance metrics, including job wait time and system throughput. Traditional heuristic-based scheduling algorithms increasingly struggle and lack the efficiency needed to meet the demands and address the complexity and scale of modern HPC systems. Consequently, recent research efforts have focused on leveraging advancements in Artificial Intelligence (AI) and Deep Learning (DL), particularly Reinforcement Learning (RL), to develop more adaptable and intelligent scheduling strategies. Previous RL-based scheduling approaches have explored a range of algorithms, from Deep Q-Networks (DQN) to Proximal Policy Optimization (PPO), and more recently, hybrid methods that integrate Graph Neural Networks (GNNs) with RL techniques. However, a common limitation across these methods is their reliance on relatively small datasets, with few methods being evaluated using large-scale, multi-million-job trace datasets representative of real-world HPC workloads. Moreover, existing RL schedulers face scalability issues due to centralized policy updates, which hinder training efficiency and performance when applied to large datasets. This study introduces a novel RL-based scheduler utilizing Decentralized Distributed Proximal Policy Optimization (DD-PPO) algorithm, which supports large-scale distributed training across multiple workers without requiring parameter synchronization at every step. By eliminating reliance on centralized updates to a shared policy, the DD-PPO scheduler enhances scalability, training efficiency, and sample utilization. Experimental validation using a large real-world dataset containing over 11.5 million job traces collected from petascale HPC systems over six years assesses the influence of dataset scale on training effectiveness and compares DD-PPO performance to traditional and advanced scheduling approaches. The experimental results demonstrate improved scheduling performance in comparison to both heuristic-based schedulers and existing RL-based scheduling algorithms.
We investigate the generation of EPR pairs between three observers in a general causally structured setting, where communication occurs via a noisy quantum broadcast channel. The most general quantum codes for this setup take the form of tripartite quantum channels. Since the receivers are constrained by causal ordering, additional temporal relationships naturally emerge between the parties. These causal constraints enforce intrinsic no-signalling conditions on any tripartite operation, ensuring that it constitutes a physically realizable quantum code for a quantum broadcast channel. We analyze these constraints and, more broadly, characterize the most general quantum codes for communication over such channels. We examine the capabilities of codes that are fully no-signalling among the three parties, positive partial transpose (PPT)-preserving, or both, and derive simple semidefinite programs to compute the achievable entanglement fidelity. We then establish a hierarchy of semidefinite programming converse bounds -- both weak and strong -- for the capacity of quantum broadcast channels for EPR pair generation, in both one-shot and asymptotic regimes. Notably, in the special case of a point-to-point channel, our strong converse bound recovers and strengthens existing results. Finally, we demonstrate how the PPT-preserving codes we develop can be leveraged to construct PPT-preserving entanglement combing schemes, and vice versa.
Emerging materials science platforms with the ability to make autonomous decisions on the fly are fundamentally changing the outlook and protocols for materials optimization and discovery. Because AI-driven self-navigating schemes can effectively reduce the total number of iterations needed to arrive at the "answer" (i.e. the best stochiometric composition for a desired physical property, optimum materials processing parameters, etc.) by significant margins, they have the potential to revolutionize materials and chemical manufacturing processes at large in research laboratory settings as well as in industrial plants. Here, we demonstrate a successful implementation of real-time closed-loop autonomous navigation of a multi-dimensional materials synthesis parameter space for fabricating phase-pure epitaxial films of a metastable phase of a functional oxide in a combinatorial pulsed laser deposition chamber. Sequential epitaxial growth iterations in search of the optimized recipe to stabilize the desired crystal phase were performed using frame-by-frame quantitative computer vision analysis of reflection high-energy electron diffraction (RHEED) images of the unit-cell level film being deposited. The autonomous scheme regularly resulted in > 30-fold reduction in the number of required experiments compared to a comprehensive mapping of the parameter space. The real-time workflow developed here can be readily extended to a variety of thin film synthesis platforms opening the door for self-driving atomic-level materials design as well as autonomous optimization of semiconductor manufacturing.
Computational Data for: "Introducing Metal-Sulfur Active Sites in Metal-Organic Frameworks via Post-Synthetic Modification for Hydrogenation Catalysis"
This repository contains the supporting data for the publication titled "Synthesis and Computational Analysis of Uranium(III)-Pnictogen Bonds" by Lauren M. Lopez, Diana Perales, Allison N. Smolek, Matthias Zeller, Bess Vlaisavljevich and Suzanne C. Bart. The manuscript associated with these data was published in J. Am. Chem. Soc. (DOI: 10.1021/jacs.5c19802). For readers only seeking the XYZ files associated with the publication, a second zipped folder called "coordinates.tgz" is included. The README.txt file describes the contents of the folder containing the input and output files.