Search NASA⌕ Search

SEARCH · Search NASA

Results for “Computer systems performance”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 757 records · Page 42

Neural Posterior Estimation for Scalable and Accurate Inverse Parameter Inference in Li-Ion Batteries

Diagnosing the internal state of Li-ion batteries is critical for battery research, operation of real-world systems, and prognostic evaluation of remaining lifetime. By using physics-based models to perform probabilistic parameter estimation via Bayesian calibration, diagnostics can account for the uncertainty due to model fitness, data noise, and the observability of any given parameter. However, Bayesian calibration in Li-ion batteries using electrochemical data is computationally intensive even when using a fast surrogate in place of physics-based models, requiring many thousands of model evaluations. A fully amortized alternative is neural posterior estimation (NPE). NPE shifts the computational burden from the parameter estimation step to data generation and model training, reducing the parameter estimation time from minutes to milliseconds, enabling real-time applications. The present work shows that NPE can infer parameters equally or more accurately than Bayesian calibration, even if it leads to higher voltage reconstruction errors. We also demonstrate that the higher computational costs for data generation are tractable even in high-dimensional cases (ranging from 6 to 27 estimated parameters). The NPE method also offers several interpretability advantages over Bayesian calibration, such as local parameter sensitivity to specific regions of the voltage curve. The NPE method is demonstrated using an experimental fast charge dataset, with parameter estimates validated against measurements of loss of lithium inventory and loss of active material. The implementation is made available in a companion repository (https://github.com/NatLabRockies/BatFIT).

25 ENERGY STORAGE↗

Computation of Auger Electron Spectra in Organic Molecules with Multiconfiguration Pair-Density Functional Theory

Efficient and accurate computation of molecular Auger electron spectra for larger systems is limited by the rapid increase in the number of doubly ionized final states as the system size grows. Here, in this work, we benchmark the application of multiconfiguration pair-density functional theory with a restricted active space (RAS) reference wave function for computing the carbon K-edge decay spectra of 20 organic molecules. Decay rates are computed within the one-center approximation. We evaluate the performance of different basis sets and on-top functionals and find that multiconfiguration pair-density functional theory achieves accuracy comparable to RAS followed by second-order perturbation theory, but at significantly lower computational cost.

Fouda, Adam E. A. [Argonne National Laboratory (AN↗

Density functional theory-based surrogate kinetic models for heterogeneous reactions of hydrocarbon intermediates on silicon carbide

The increasing demand for high-performance materials in advanced technologies highlights the importance of achieving a fundamental understanding and potential control of silicon carbide (SiC) deposition processes. However, existing models often lack sufficient theoretical detail, relying heavily on empirical data and offering limited predictive capability. In particular, the complex surface chemistry governing SiC growth remains poorly understood. This study addresses these challenges by employing density functional theory (DFT) to investigate key heterogeneous reactions involving hydrocarbon intermediates on SiC surfaces, including dehydrogenation, hydrogenation, and carbon deposition. Transition state searches were conducted to identify reaction pathways and energy barriers. While first-principles calculations offer high accuracy, they are computationally intensive. To extend the utility of these first-principles results, vibrational analyses were performed using phonon-based statistical thermochemistry to compute temperature-dependent reaction rates which were used to develop Arrhenius-type surrogate kinetic models. Furthermore, the resulting framework provides a more rigorous, physically grounded basis for integrating atomistic insights into continuum-scale modeling, ultimately enabling improved prediction and optimization of SiC film growth in high-performance material systems.

Density Functional Theory↗

GPU acceleration of hybrid functional calculations in the SPARC electronic structure code

We present a Graphics Processing Unit (GPU)-accelerated version of the real-space SPARC electronic structure code for performing hybrid functional calculations in generalized Kohn–Sham density functional theory. In particular, we develop a batch variant of the recently formulated Kronecker product-based linear solver for the simultaneous solution of multiple linear systems. We then develop a modular, math kernel based implementation for hybrid functionals on NVIDIA architectures, where computationally intensive operations are offloaded to the GPUs, while the remaining workload is handled by the central processing units (CPUs). Considering bulk and slab examples, we demonstrate that GPUs enable up to 8× speedup in node-hours and 80× in core-hours compared to CPU-only execution, reducing the time to solution on V100 GPUs to around 300 s for a metallic system with over 6000 electrons, and significantly reducing the computational resources required for a given wall time.

Kohn-Sham density functional theory↗

Computational Analysis of Hydraulic Efficiency of Michigan DOT Cover C

Drainage structures are used to capture stormwater runoff in streets and highways in urban environments. These drainage structures, which typically consist of catch basins with grates, inlets, or combination grates/inlets, collect stormwater runoff and discharge through buried conveyance systems. They are strategically placed for public safety in curb and gutter systems to provide efficient drainage of water from roadways and thus reduce the risk of hydroplaning. The performance of drainage structures is measured in terms of hydraulic efficiency, which is defined as the percentage of flow captured by the basin as compared to the total flow drainage to the structure. Understanding of the performance of these drainage structures helps designers properly space inlets to promote an economic design that ensures the safety of the traveling public. The current design methodology used to determine drainage structure follows guidelines established in the current Michigan Department of Transportation (MDOT) Drainage Manual (2006). The guidelines in the MDOT Drainage Manual were modeled after the Federal Highway Administration’s (FHWA) Hydraulic Engineering Circular 22 “Urban Drainage Design” (HEC-22). HEC-22 includes empirically derived equations to calculate the interception capacity of drainage structures for several commonly used grate configurations, such as the parallel bar, curved vane and tilt bar grates, which are based on a research study performed by Burgi et al. in the 1970s. MDOT uses several drainage structures to capture runoff that are detailed as Standard Plans. Many of these drainage structures utilize sinusoidal type grates that are not described in HEC-22. Physical modeling of these structures has been limited, posing the need to have them analyzed to verify their capture efficiency. Current MDOT practice is to assume a similar sized reticuline grate, as described in HEC-22, for capture efficiencies. Until recently, evaluating the hydraulic performance of drainage structures was limited to physical modeling in a hydraulics laboratory. With advances in engineering software and computing power, computational fluid dynamics (CFD) modeling has become a more cost-effective alternative. The Federal Highway Administration (FHWA) provides states the option to evaluate their drainage structures using CFD through the Transportation Pooled Fund Program. This study, “Computational Analysis of Hydraulic Efficiency of Michigan DOT Cover C,” was carried out using the pooled fund. MDOT’s Cover C was chosen as the first test candidate, given its similar sinusoidal pattern to other MDOT grates, but it is typically used for high-volume, higher speed applications. A similar version, Cover CX, is used on interstate highways but does not have traverse bars for bicycle safety. Additional grates may be considered for evaluation in the future.

42 ENGINEERING↗

Towards a Deeper Fundamental Understanding of (Al,Sc)N Ferroelectric Nitrides

Density functional theory (DFT) calculations, within the virtual crystal alloy approximation, are performed, along with the development of a Landau-type model employing a symmetry-allowed analytical expression of the internal energy and having parameters determined from first principles, to investigate properties and energetics of Al1-xScxN ferroelectric nitrides in their hexagonal forms. These DFT computations and this model predict the existence of two different types of minima, namely, the fourfold-coordinated wurtzite (WZ) polar structure and a five-fold coordinated paraelectric hexagonal phase (denoted as H5), for any Sc composition up to 40%. The H5 minimum progressively becomes the lowest-energy state within hexagonal symmetry as the Sc concentration increases from 0 to 0.4. Furthermore, the model points to several key findings. Examples include the crucial role of the coupling between polarization and strains to create the WZ minimum, in addition to polar and elastic energies, and that the origin of the H5 state overcoming the WZ phase as the global minimum within hexagonal symmetry when increasing the Sc composition mostly lies in the compositional dependency of only two parameters-one linked to the polarization and another one being purely elastic in nature. Other examples are that forcing Al1-xScxN systems to have no or a weak change in lattice parameters when heating them allows us to reproduce their finite-temperature polar properties well and that a value of the axial ratio close to that of the ideal WZ structure implies a large polarization at low temperatures but not necessarily at high temperatures because of the ordered-disordered character of the temperature-induced formation of the WZ state. Such findings should allow for a better fundamental understanding of (Al,Sc)N ferroelectric nitrides, which may be used to design efficient devices having, e.g., low operating voltages.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Latent space dynamics identification for interface tracking with application to shock-induced pore collapse

Capturing sharp, evolving interfaces remains a central challenge in reduced-order modeling, especially when data is limited and the system exhibits localized nonlinearities or discontinuities. Here, we propose LaSDI-IT (Latent Space Dynamics Identification for Interface Tracking), a data-driven framework that combines low-dimensional latent dynamics learning with explicit interface-aware encoding to enable accurate and efficient modeling of physical systems involving moving material boundaries. At the core of LaSDI-IT is a revised autoencoder architecture that jointly reconstructs the physical field and an indicator function representing material regions or phases, allowing the model to track complex interface evolution without requiring detailed physical models or mesh adaptation. The latent dynamics are learned through linear regression in the encoded space and generalized across parameter regimes using Gaussian process interpolation with greedy sampling. We demonstrate LaSDI-IT on the problem of shock-induced pore collapse in high explosives, a process characterized by sharp temperature gradients and dynamically deforming pore geometries. The method achieves relative prediction errors below 9% across the parameter space, accurately recovers key quantities of interest such as pore area and hot spot formation, and matches the performance of dense training with only half the data. This latent dynamics prediction was 10 6 times faster than the conventional high-fidelity simulation, proving its utility for multi-query applications. These results highlight LaSDI-IT as a general, data-efficient framework for modeling discontinuity-rich systems in computational physics, with potential applications in multiphase flows, fracture mechanics, and phase change problems.

Gaussian process↗

Sachdev-Ye-Kitaev model on a noisy quantum computer

Here we study the SYK model -- an important toy model for quantum gravity on IBM's superconducting qubit quantum computers. By using a graph-coloring algorithm to minimize the number of commuting clusters of terms in the qubitized Hamiltonian, we find the gate complexity of the time evolution using the first-order product formula for N Majorana fermions is $\mathscr{O}$(N 5 J 2 t 2 /ε) where J is the dimensionful coupling parameter, t is the evolution time, and ε is the desired precision. With this improved resource requirement, we perform the time evolution for N=6,8 with maximum two-qubit circuit depth of 343. We perform different error mitigation schemes on the noisy hardware results and find good agreement with the exact diagonalization results on classical computers and noiseless simulators. In particular, we compute return probability after time t and out-of-time order correlators (OTOC) which is a standard observable of quantifying the chaotic nature of quantum systems.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Single-Photon Generation: Materials, Techniques, and the Rydberg Exciton Frontier

Due to their quantum nature, single-photon emitters (SPE) generate individual photons in bursts or streams. They are paramount in emerging quantum technologies such as quantum key distribution, quantum repeaters, and measurement-based quantum computing. Many such systems have been reported in the last three decades, from rubidium atoms coupled to cavities to semiconductor quantum dots and color centers implanted in waveguides. This review article highlights different solid-state and atomic systems with on-demand and controlled single-photon generation. We discuss and compare the performance metrics, such as purity and indistinguishability, for these sources and evaluate their potential for different applications. Finally, a new potential single-photon source, based on the Rydberg exciton in solid-state metal oxide thin films, is introduced, where we discuss its promising features and unique advantages in fabricating quantum chips for quantum photonic applications.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Influence of Process Parameter and Build Rate Variations on Defect Formation in Laser Powder Bed Fusion SS316L

Laser powder bed fusion (LPBF) is an additive manufacturing process that has gained interest for its material fabrication due to multiple advantages, such as the ability to print parts with small feature sizes, good mechanical properties, reduced material waste, etc. However, variations in the key process parameters in LPBF may result in the instantiation of porosity defects and variation in build rate. Particularly, volumetric energy density (VED) is a variable that encapsulates a number of those parameters and represents the amount of energy input from the laser source to the feedstock. VED has been traditionally used to inform the quality of the printed part but different values of VED are presented as optimal values for certain material systems. An optimal VED value can be maintained by changing the key process parameters so that various combinations yield a constant value. In this study, an optimal constant VED value is maintained while printing SS316L with variable key processing parameters. Porosity analysis is performed using optical microscopy, as well as X-ray computed tomography, to reveal the volume density and distribution of those pores. Two primary defect categories are identified, namely lack of fusion and porosity induced by balling defects. The findings indicate that, even at optimal VED, variations in process parameters can significantly influence defect type, underscoring the sensitivity of defect formation to the variation of these parameters. Furthermore, a minor change in the build rate, driven by adjustments in process parameters, was found to influence defect categories. These findings emphasize that fine tuning the process parameters and build rate is essential to minimize defects. Finally, fiducial marks have been identified as a source of unintentional porosity defects. These results enable the refinement of process parameters, ultimately optimizing LPBF to achieve enhanced material density and expedite the printing.

36 MATERIALS SCIENCE↗

System Advisor Model (SAM) [Slides]

The System Advisor Model (SAM) is a free software that enables detailed performance and financial analysis for energy systems. • Conducts technoeconomic analysis of energy technologies to facilitate planning and decision making.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Dissociation of gas-phase anisole induced by low-energy electron interactions: understanding patterns of aromatic bond cleavage

Abstract In view of elucidating the fragmentation patterns of aromatic systems induced by low-energy electron interactions, dissociative electron attachment (DEA) to gas-phase anisole was performed. Anionic fragments resulting from this DEA process were detected by a quadrupole mass spectrometer, and ion yields of those fragments as a function of incident electron energy were rendered. Our study showed the formation of CH 3 − , HCC − , and OCH 3 − fragments, suggesting that various dissociation channels proceed out of DEA to anisole. We employed density functional theory to compute thermodynamic threshold energies for each potential dissociation channel. Those theoretical calculations supported the prediction that the CH 3 − and OCH 3 − fragments form via mechanisms of single-bond cleavage; the HCC − fragments may form through two-, three-, or four-body dissociation channels that entail hydrogen transfers and the cleavage of multiple aromatic bonds. The experimental resonance energies that form the CH 3 − , HCC − , and OCH 3 − fragments were 6.0 eV, 5.8 and 9.7 eV, and 9.8 eV, respectively. Given the classification of anisole as a monosubstituted aromatic species, our results explain generalizable patterns of electron-mediated dissociation in aromatic systems.

Finley, Jacob (ORCID:0009000606963352)↗

Optimizing qubit control pulses for state preparation

In the burgeoning field of quantum computing, the precise design and optimization of quantum pulses are essential for enhancing qubit operation fidelity. This study focuses on refining the pulse engineering techniques for superconducting qubits, employing a detailed analysis of square and Gaussian pulse envelopes under various approximation schemes. We evaluated the effects of coherent errors induced by naive pulse designs. Furthermore, we identified the sources of these errors in the Hamiltonian model’s approximation level. We mitigated these errors through adjustments to the external driving frequency and pulse durations, thus implementing a pulse scheme with stroboscopic error reduction. Our results demonstrate that these refined pulse strategies improve performance and reduce coherent errors. Moreover, the techniques developed herein are applicable across different quantum architectures, such as ion-trap, atomic, and photonic systems.

Chirp modulation↗

Optical Control of Adaptive Nanoscale Domain Networks

Adaptive networks can sense and adjust to dynamic environments to optimize their performance. Understanding their nanoscale responses to external stimuli is essential for applications in nanodevices and neuromorphic computing. However, it is challenging to image such responses on the nanoscale with crystallographic sensitivity. Here, the evolution of nanodomain networks in (PbTiO 3 ) n /(SrTiO 3 ) n superlattices (SLs) is directly visualized in real space as the system adapts to ultrafast repetitive optical excitations that emulate controlled neural inputs. The adaptive response allows the system to explore a wealth of metastable states that are previously inaccessible. Their reconfiguration and competition are quantitatively measured by scanning x-ray nanodiffraction as a function of the number of applied pulses, in which crystallographic characteristics are quantitatively assessed by assorted diffraction patterns using unsupervised machine-learning methods. The corresponding domain boundaries and their connectivity are drastically altered by light, holding promise for light-programable nanocircuits in analogy to neuroplasticity. Phase-field simulations elucidate that the reconfiguration of the domain networks is a result of the interplay between photocarriers and transient lattice temperature. The demonstrated optical control scheme and the uncovered nanoscopic insights open opportunities for the remote control of adaptive nanoscale domain networks.

36 MATERIALS SCIENCE↗

Dielectric-Engineered Monolayer MoS 2 Memtransistors for Brain-Inspired Computing with High Recognition Accuracy

Two-dimensional transition metal dichalcogenides (2D-TMDs)-based memtransistors have emerged as promising candidates for neuromorphic hardware due to their exceptional ability to emulate synaptic behavior. However, many existing 2D-TMDs memtransistors rely on polycrystalline channels with grain boundaries or defects introduced through postgrowth treatments, raising concerns about material integrity and the preservation of intrinsic properties. Here, in this work, we demonstrate a monocrystalline monolayer MoS 2 memtransistor fabricated on a silicon nitride (SiN X ) substrate, achieving a large resistive switching ratio of 10 4 , a dynamic range exceeding 90, along with highly linear and symmetric weight updates, minimal cycle-to-cycle variability, and low device-to-device variability. These attributes are critical for enabling high-performance neuromorphic hardware. Based on experimental data, we further show that these artificial synapses enable a recognition accuracy of more than 97% on the MNIST handwritten digits data set. Our findings present a straightforward approach to realizing 2D-TMDs memtransistors through dielectric engineering, offering a promising platform for next-generation neuromorphic computing systems.

2D TMDs↗

Architecting the Third Dimension of Electrochemical Energy Storage

Three-dimensional (3D) architectural design has emerged as a powerful strategy to push electrochemical energy storage (EES) devices beyond the intrinsic limitations of conventional two-dimensional (2D) electrodes. While planar architectures enable high packing density and mature manufacturing, they suffer from limited ion transport and low active-material loading. In contrast, 3D architectures introduce low-tortuosity networks and high surface area that enhance charge and mass transport while supporting thick, high mass-loading electrodes. However, their practicality remains hindered by challenges in volumetric density, mechanical stability, and large-scale manufacturability. Here, this Perspective examines the key evaluation and design principles that govern 3D device performance. We discuss the fundamental trade-offs between porosity, volumetric density, and mechanical stability that shape 3D design and highlight emerging strategies for integrating materials engineering, structural optimization, device integration, computational modeling, and scalable manufacturing. By aligning structural functionality with manufacturability, 3D architectures can evolve from laboratory prototypes to commercially viable energy storage systems.

25 ENERGY STORAGE↗

High-Speed and High-Quality Field Welding Repair Based on Advanced Non-Destructive Evaluation and Numerical Modeling

Creep strength-enhanced ferritic (CSEF) steels such as Grade 91 (9Cr-1Mo-V) and Grade 92 (Fe-9Cr-2W-0.5Mo) steels are widely used in the fossil-fuel-fired and nuclear power plants. The weld integrity of these steels is crucial for power plants' safe and reliable operations. Due to harsh service conditions, the steel weld can become susceptible to environmental degradation. Field welding repair is used to restore the degraded weld’s performance where a controlled temper-bead welding technique is commonly used to temper the freshly formed martensite during welding. However, knowledge of weld repairability is limited and experimental trial and error optimization to achieve desired microstructure and joint properties is expensive and time-consuming. Many existing computational models, e.g., finite element models, are limited to solving heat conduction equation and ignoring convective heat transfer due to molten metal flow. These models can result in over-prediction of peak temperatures of weld pool and heat-affected zone (HAZ), which in turn can affect the accuracy of tempering prediction. Moreover, these finite element models require an input of the deposit profiles in advance and thus limits the usability of these models. Here, a molten pool-based, multi-pass multi-layer model has been developed based on computational fluid dynamics (CFD) approach with the Volume of Fluid (VOF) method. The model calculates the bead formation, thereby eliminating the need for pre-determined bead profiles required by finite element models. For computational efficiency, a coordinate system attached to the moving heat source is utilized. A subroutine is developed to convert the temperature profiles in the reference frame stationary to the heat source to that stationary to the workpiece. The converted thermal cycles are then imported into a microstructure model to compute the tempering kinetics and resultant hardness using a Johnson-Mehl-Avrami-Kolmogorov (JMAK), and modified Grange-Baughman parameter. The modeling approach is first developed and validated on single- and multi-pass deposition of stainless steel filler metal onto a SA-533 high strength steel substrate. The models are then applied to a multi-pass V-groove repair weld of Grade 91 steel plate as well as directed energy deposition of Grade 92 steel. Non-destructive characterization of microstructures was performed on Grade 91 and 92 steel welds. Two welding processes, cold metal transfer (CMT) and flux-cored arc welding (FCAW), were investigated for the Grade 91 steel weld samples. For the Grade 92 weld samples, three different heat inputs (low, medium, and high) of gas tungsten arc welding (GTAW) were utilized to replicate traditional field welding processes. The non-destructive evaluation (NDE) method used for this research was immersion ultrasonic testing (UT) using a micro-resolution ultrasonic imaging methodology specifically designed to operate in the through-transmission configuration operating at 20 MHz of frequency. The system used a focused ultrasonic beam spot size diameter between 250-300 μm, and a 6 μm laser vibrometer spot size for detection, to produce highly defined images with longitudinal and mode-converted shear waves. From the micro-resolution ultrasonic C-scan images, three microstructural regions, i.e., weld metal (WM), HAZ, and base metal (BM), were clearly identifiable. Various levels of ultrasonic amplitudes distributed over the three regions were correlated with electron beam backscattered diffraction (EBSD) images using grain size, grain boundaries, and dislocation densities. The results showed that areas with relatively higher ultrasonic amplitude levels were associated with smaller grains and higher dislocation densities, while areas with lower amplitude levels were associated with larger grains and lower dislocation densities. In addition, ultrasonic velocity data obtained across the three different weld microstructural regions of Grade 91 test samples were correlated with optical metallographic images and hardness measurements. The results showed distinctive decreases in ultrasonic velocity and hardness over the HAZ region, where weld failures often occur during service.

36 MATERIALS SCIENCE↗

Synchronization for CXL Based Memory

Compute Express Link (CXL) is an important emerging standard for disaggregated memory. While this standard provisions coherency across numerous hosts and devices, implementing hardware support for type three devices is challenging. In this work, we look at the overhead of software synchronization and using software-based coherency. Moreover, we discuss the limits of software-based coherency in fully expressing modern synchronization techniques for a CXL-based disaggregate memory system. We demonstrate our approach using a CXL hardware prototype and running a version of the famous Peterson Lock (enhanced to run with more than two threads). We analyze its performance and share how more advanced synchronization techniques might interact with software-based coherence CXL hardware and program execution models.

High Performance Computing (HPC)↗