Search NASA⌕ Search

SEARCH · Search NASA

Results for “Computer architecture”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 523 records · Page 29

Swap Path Network for Robust Person Search Pre-training

This code corresponds to the WACV25 conference paper, "Swap Path Network for Robust Person Search Pre-training". In that paper, we introduce a new model for the person search task called the Swap Path Net (SPNet). The person search task is a problem in computer vision, where we locate and rank matches to an image of a query person in a set of other images where we want to find them. We also introduce a novel pre-training algorithm specific to the Swap Path Net architecture. The code implements pre-training and fine-tuning of the Swap Path Net (SPNet). This includes ingesting image datasets and updating the weights of the SPNet neural network to train it for the person search task. The repository contains code, configs, and instructions to reproduce all results from the paper.

Jaffe, LucasW [Lawrence Livermore National Laborat↗

Toward a persistent event-streaming system for high-performance computing applications

High-performance computing (HPC) applications have traditionally relied on parallel file systems and file transfer services to manage data movement and storage. Alternative approaches have been proposed that use direct communications between application components, trading persistence and fault tolerance for speed. Event-driven architectures, as popularized in enterprise contexts, present a compelling middle ground, avoiding the performance cost and API constraints of parallel file systems while retaining persistence and offering impedance matching between application components. However, adapting streaming frameworks to HPC workloads requires addressing challenges unique to HPC systems. This paper investigates the potential for a streaming framework designed for HPC infrastructures and use cases. We introduce Mofka, a persistent event-streaming framework designed specifically for HPC environments. Mofka combines the capabilities of a traditional streaming service with optimizations tailored to the HPC context, such as support for massively multicore nodes, efficient scaling for large producer-consumer workflows, RDMA-enabled high-performance network communications, specialized network fabrics with multiple links per node, and efficient handling of large scientific data payloads. Built using the Mochi suite of HPC data service components, Mofka provides a lightweight, modular, and high-performance solution for persistent streaming in HPC systems. We present the architecture of Mofka and evaluate its performance against Kafka and Redpanda using benchmarks on diverse platforms, including Argonne's Polaris and Oak Ridge's Frontier supercomputers, showing up to 8× improvement in throughput in some scenarios. We then demonstrate its utility in several real-world applications: a tomographic reconstruction pipeline, a workflow for the discovery of metal-organic frameworks for carbon capture, and the instrumentation of Dask workflows for provenance tracking and performance analysis.

HPC↗

Understanding and design of interstitial oxygen conductors

Highly efficient oxygen-active materials that react with, absorb, and transport oxygen is essential for fuel cells, electrolyzers and related applications. While vacancy-mediated oxygen-ion conductors have long been the focus of research, they are limited by high migration barriers at intermediate temperatures (400–600 °C), which hinder their practical applications. In contrast, interstitial oxygen conductors exhibit significantly lower migration barriers enabling higher ionic conductivity at lower temperatures. This review systematically examines both well-established and recently identified families of interstitial oxygen-ion conductors, focusing on how their unique structural motifs such as corner-sharing polyhedral frameworks, isolated polyhedral, and cage-like architectures, facilitate low migration barriers through interstitial and/or interstitialcy diffusion mechanisms. A central discussion of this review focuses on the evolution of design strategies, from targeted donor doping, element screening, to physical-intuition descriptor material screening and machine learning approach, which leverage computational tools to explore vast chemical spaces in search for new interstitial conductors. The success of these strategies demonstrates that a significant, largely unexplored space remains for discovering high-performing interstitial oxygen conductors. Crucial features enabling high-performance interstitial oxygen diffusion include the availability of electrons for oxygen reduction and sufficient structural flexibility with accessible volume for interstitial accommodation and migration. This review concludes with a forward-looking perspective, proposing a knowledge-driven methodology that integrates current understanding with data-centric approaches to identify promising interstitial oxygen conductors outside traditional search paradigms. These approaches are expected to significantly accelerate the development of high-performance interstitial oxygen conductors for a variety of oxygen-active applications, ultimately paving the way for more efficient and sustainable energy technologies.

Interstitial oxygen conductors↗

Dynamical Signatures of Thermotoga maritima Maltose-Binding Proteins Affected by Ligand Binding

Functional segregation among protein isoforms depends on the interplay of their overall structures and the molecular dynamics of these structures. Thermotoga maritima maltose-binding protein (tmMBP) isoforms show size-dependent differential binding of maltose and malto-oligomers while maintaining remarkable fold conservation. This differential behavior needs detailed characterization in native-like aqueous conditions to understand the effects of protein dynamics on ligand binding and recognition. Small-angle neutron scattering (SANS), neutron spin echo (NSE) spectroscopy, and dynamic light scattering (DLS) were used in conjunction with previously published computational molecular dynamics (MD) simulations to understand the dynamic behavior of tmMBPs experimentally. SANS provided information on the overall structure of the molecules, while NSE was used to determine the dynamics in the nanosecond time scale. Both tmMBP2 and tmMBP3 have a bidomain architecture linked with a flexible hinge, with the binding pocket sitting in the cleft between the two domains. tmMBP2 and tmMBP3 showed different solution dynamics, with the translational and rotational components dominating the dynamics of both systems, resulting in a clear differentiation of their diffusion pattern. A faster dynamics component was also observed and was attributed to segmental dynamics. Differences observed between the ligand-free (apo) and ligand-bound (holo) states of the two proteins are attributed to conformational entropy. Our results highlight the intricacies of how structure and dynamics can together shape binding to a repertoire of substrates in structurally similar proteins.

Diffusion↗

VISION: a modular AI assistant for natural human-instrument interaction at scientific user facilities

Scientific user facilities, such as synchrotron beamlines, are equipped with a wide array of hardware and software tools that require a codebase for human-computer-interaction. This often necessitates developers to be involved to establish connection between users/researchers and the complex instrumentation. The advent of generative AI presents an opportunity to bridge this knowledge gap, enabling seamless communication and efficient experimental workflows. Here we present a modular architecture for the Virtual Scientific Companion by assembling multiple AI-enabled cognitive blocks that each scaffolds large language models (LLMs) for a specialized task. With VISION, we performed LLM-based operation on the beamline workstation with low latency and demonstrated the first voice-controlled experiment at an x-ray scattering beamline. The modular and scalable architecture allows for easy adaptation to new instruments and capabilities. Development on natural language-based scientific experimentation is a building block for an impending future where a science exocortex—a synthetic extension to the cognition of scientists—may radically transform scientific practice and discovery.

36 MATERIALS SCIENCE↗

Error mitigation, optimization, and extrapolation on a trapped-ion testbed

Current noisy intermediate-scale quantum (NISQ) trapped-ion devices are subject to errors which can significantly impact the accuracy of calculations if left unchecked. A form of error mitigation called zero noise extrapolation (ZNE) can decrease an algorithm’s sensitivity to these errors without increasing the number of required qubits. Here we explore different methods for integrating this error mitigation technique into the Variational Quantum Eigensolver (VQE) algorithm for calculating the ground state of the HeH + molecule at 0.8 Å in the presence of experimental noise. Using the Quantum Scientific Computing Open User Testbed (QSCOUT) trapped-ion device, we test three methods of scaling noise for extrapolation: time stretching the two-qubit gates, scaling the sideband detuning parameter, and inserting two-qubit gate identity operations into the ansatz circuit. We find that time stretching and sideband detuning scaling fail to scale the noise on our particular hardware in a way that can be extrapolated to zero noise. Scaling our noise with global gate identity insertions and extrapolating after variational optimization, we achieve error suppression of 96.8%, resulting in an energy estimate within –0.004 ± 0.04 hartree of the ground state energy. This is an improvement, but still outside the chemical accuracy threshold of 0.0016 hartree. Furthermore, our results show that the efficacy of this error mitigation technique depends on choosing the correct implementation for a given device architecture.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Ultra-thick three-dimensional interpenetrating graphene electrode architectures for high volumetric density energy storage

For electrochemical energy storage, increasing the electrode thickness is an effective approach to achieving higher energy density from a given material. However, this often compromises ion transport, leading to diminished performance. Here, in this study, we present a novel platform for fabricating complex 3D interpenetrating electrode structures via photo-polymerization 3D printing, integrated with computational structural optimization for energy storage. The platform employs an acrylate resin system infused with graphene oxide (GO), enabling high-fidelity printing of optimized porous structures and facilitating efficient electron and ion transport in ultra-thick electrodes. The optimized 3D layouts substantially enhance energy and power densities compared to conventional configurations, ensuring superior material utilization and minimal ohmic losses. Supercapacitors fabricated using this approach achieved an exceptional energy density of 4.7 Wh L−1 at a power density of 1689.0 W L−1, surpassing traditional designs. This work underscores the transformative role of structural optimization in advancing electrochemical performance and establishes a versatile pathway for developing next-generation energy storage systems with exceptional efficiency and functionality.

Wang, Zhen [University of California, Berkeley, CA↗

Performance evaluations of signed and unsigned noisy approximate quantum Fourier arithmetic

The Quantum Fourier Transform (QFT) grants competitive advantages, especially in resource usage and circuit approximation, for performing arithmetic operations on quantum computers, and offers a potential route toward a numerical quantum-computational paradigm. In this paper, we utilize efficient techniques to implement QFT-based integer addition and multiplications. These operations are fundamental to various quantum applications including Shor’s algorithm, weighted-sum optimization problems in data processing and machine learning, and quantum algorithms requiring inner products. We carry out performance evaluations of these implementations based on IBM’s superconducting-qubit architecture using different compatible noise models. We isolate the sensitivity of the component quantum circuits on both one-/two-qubit gate error rates, and the number of the arithmetic operands’ superposed integer states. We analyze performance and identify the most effective approximation depths for unsigned quantum addition and quantum multiplication within the given context. We then perform a similar analysis of signed addition and compare to the unsigned results. We observe significant dependency of the optimal approximation depth on the degree of machine noise and the number of superposed states in certain performance regimes. Finally, we elaborate on the algorithmic challenges—relevant to signed, unsigned, modular and non-modular versions—that could also be applied to current implementations of QFT-based subtraction, division, exponentiation, and their potential tensor extensions. Here, we analyze the performance trends in our results and speculate on possible future developments within this computational paradigm.

Computational models↗

Demonstrating the data center as a flexible grid asset using a C-HIL setup

Increasing data center demand is outpacing grid infrastructure development. Artificial intelligence workloads and hyperscale cloud growth are creating unprecedented demand for power, while traditional grid expansion faces multiyear development timelines. Verrus is developing an innovative datacenter solution for this challenge, data centers that act as active grid-supportive assets rather than passive loads. Our approach integrates a novel grid-aware power flow management system with battery energy storage systems(BESS) into a microgrid-controlled, medium-voltage power distribution architecture that delivers critical capabilities, such as: * Fast response to grid disturbances such over/ under voltage or over/ under frequency * Demand flexibility that can service requests from the utility within 10 s * Uninterrupted transition to islanded operation during grid outages * Continuous uptime assurance for compute loads while maintaining all customer service level agreements. Through Verrus' strategic partnership with the National Renewable Energy Laboratory (NREL), these capabilities were validated using NREL's Advanced Research on Integrated Energy Systems (ARIES) virtual emulation environment to model a 70-MW grid-interactive data center. This paper outlines the design, methodology, and results of this emulated deployment, demonstrating that data centers can provide both critical load resilience and ancillary grid support without compromising uptime requirements. Specifically, we present a digital real time simulation of a 70 MW data center integrated with a physical microgrid controller, and demonstrate the data center response in the event of a grid voltage and frequency event, utility demand response request and utility outage.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Dynamic crushing of metal lattice metamaterials: Shock mode diagrams and transition to topology-independent compaction regime

Additively manufactured lattice metamaterials offer design versatility in strength and energy absorption and provide an additional degree of freedom through the selection of the lattice topology. Under quasistatic loading, the unit cell structure can strongly affect the stiffness, yield, and post-yield behavior, but whether and to what degree the effect of lattice topology persists into dynamic loading scenarios, up to the compaction shock regime, has not been established. LLNL’ s ALE3D hydrocode was used to perform a computational investigation of dynamic loading in multiple lattice types, including the gyroid, octet, Schwarz D, and rhombic dodecahedron, under impact velocities from 0.25 to 2.25 km/s. Shock Hugoniots for each lattice topology are generated and compared, suggesting that above a critical velocity, distinctions between architectures may not persevere and compacted lattices behave similarly. Here, to investigate the transition between topology-dependent quasistatic compression and the topology-independent regime above the critical velocity, a one-dimensional elastic-linear hardening plasticity-densified solid (E-LHP-DS) shock model for lattice materials was developed that relies upon confined compression to link the quasistatic and shock mechanics. Unlike similar works, the model does not assume rigid behavior prior to yield or locking behavior at densification, allowing a richer exploration of lattice mechanics. With only six parameters, the analytical model simultaneously fit quasistatic confined compression simulations for relative densities 0.1 $≤ \bar{ρ} ≤$ 0.9 and predicted dynamic compaction behavior to traverse several distinct shock modes, each defined by a critical impact speed (equivalently, critical stresses). Comparing the numerical results to the one-dimensional E-LHP-DS shock model predictions suggests that the topology-independence under strong shocks is linked to the onset of densification, which can be predicted based on quasistatic confined compression results.

Cellular material↗

Bayesian Optimized Deep Ensemble for Uncertainty Quantification of Deep Neural Networks: a System Safety Case Study on Sodium Fast Reactor Thermal Stratification Modeling

Deep neural networks (DNNs) are increasingly important to scientific computing and engineering system simulations. Accurate uncertainty quantification (UQ) for DNNs is critical in safety-sensitive engineering domains. Traditional Deep Ensemble (DE) methods, while easy to implement, frequently suffer from poorly calibrated uncertainty estimates and limited predictive accuracy due to reliance on fixed architectures with varied weight initializations. To address these issues, we introduce a workflow that combines Bayesian Optimization (BO) and DE. The workflow is modular, scalable, and integrates parallel BO initialized with Sobol sequences to individually optimize the hyperparameters of each ensemble member. This method enhances ensemble diversity, improves predictive accuracy, and provides reliable uncertainty estimates. We evaluate the proposed BODE approach in a sodium fast reactor thermal stratification modeling case study, where we used a densely connected convolutional neural network to predict turbulent viscosity during the reactor transient with consideration of data noise. We benchmark its performance against several optimization approaches, including baseline deep ensemble, evolutionary algorithm-optimized ensemble, ensemble formed via random search combined with greedy selection, and a BO ensemble using random initialization. Here, our results demonstrate superior performance of the developed BODE approach. In noise-free scenarios, BODE notably reduces incorrect aleatoric uncertainty and significantly enhances predictive accuracy. Under conditions of 5% and 10% Gaussian noise, BODE adaptively quantifies uncertainty proportional to data noise, achieving up to an 80% reduction in root mean square error compared to baseline methods and producing well-calibrated prediction intervals.

Bayesian optimization↗

Building Qudit-Based Quantum Computing Processors using Superconducting RF Cavities

Superconducting radio frequency (SRF) cavities provide an excellent platform for storing quantum information as quantum d-level systems (qudits) due to their exceptionally long lifetimes and large accessible Hilbert spaces. A common strategy to manipulate the states is to use a nonlinear element like a transmon. There are, however, several challenges to building a 3D SRF architecture while maintaining a long cavity lifetime. We demonstrate our successful integration of transmons with single-cell Nb SRF cavities and the ability to prepare several non-classical states. Finally, we discuss our strategies to improve the coherence times, gate schemes, and extend the system for building a multi-qudit quantum processor.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

EV SALaD 2023 Demonstration: Best Practices and Mitigations for Protecting EVSE Infrastructure

The Electric Vehicle Secure Architecture Laboratory Demonstration (EV SALaD) program is a demonstration of cybersecurity best practices for high-power electric vehicle (EV) charging infrastructure led by Idaho National Laboratory (INL), in collaboration with other DOE National Laboratories participating in the EVs at Scale Consortium.a Sandia National Laboratories (SNL) and Pacific Northwest National Laboratory (PNNL) participated in the first 2-year (FY22-23) demonstration cycle for EV SALaD. This report documents the FY23 demonstration, the second in a series of demonstrations and collaborations in deploying and operating cybersecure EV charging infrastructure. It includes a summary of improvements from the FY22 demonstration, technical analysis of the FY23 demonstration, how the research demonstrates cyber-physical and cybersecurity best practices for high-power EV charging infrastructure, and related impacts to national and energy security. For EV SALaD, the FY22 demonstration focused on the detection, ranking, and prioritization of anomalous events for high-power EV charging. The FY23 demonstration additionally included the demonstration of cybersecurity best practices, which included protection and mitigation solutions to prevent, respond, and recover from anomalous events. During the demonstrations, the multi-lab EV SALaD team conducted a Test Effect Payload (TEP)b evaluation on extreme fast charger (XFC) hardware equipped with Cerberus, a detection and response solution, to demonstrate anomaly detection and mitigation cybersecurity best practices against cyber-enabled events.

33 ADVANCED PROPULSION SYSTEMS↗

Fourier-MIONet: Fourier-enhanced multiple-input neural operators for multiphase modeling of geological carbon sequestration

Geologic carbon sequestration (GCS) is a safety-critical technology that aims to reduce the amount of carbon dioxide in the atmosphere, which also places high demands on reliability. Multiphase flow in porous media is essential to understand CO 2 migration and pressure fields in the subsurface associated with GCS. However, numerical simulation for such problems in 4D is computationally challenging and expensive, due to the multiphysics and multiscale nature of the highly nonlinear governing partial differential equations (PDEs). It prevents us from considering multiple subsurface scenarios and conducting real-time optimization. Here, we develop a Fourier-enhanced multiple-input neural operator (Fourier-MIONet) to learn the solution operator of the problem of multiphase flow in porous media. Fourier-MIONet utilizes the recently developed framework of the multiple-input deep neural operators (MIONet) and incorporates the Fourier neural operator (FNO) in the network architecture. Once Fourier-MIONet is trained, it can predict the evolution of saturation and pressure of the multiphase flow under various reservoir conditions, such as permeability and porosity heterogeneity, anisotropy, injection configurations, and multiphase flow properties. Compared to the enhanced FNO (U-FNO), the proposed Fourier-MIONet has 90% fewer unknown parameters, and it can be trained in significantly less time (about 3.5 times faster) with much lower CPU memory (<15%) and GPU memory (<35%) requirements, to achieve similar prediction accuracy. In addition to the lower computational cost, Fourier-MIONet can be trained with only 6 snapshots of time to predict the PDE solutions for 30 years. Furthermore, we observed that Fourier-MIONet can maintain good accuracy when predicting out-of-distribution (OOD) data. The excellent generalizability of Fourier-MIONet is enabled by its adherence to the physical principle that the solution to a PDE is continuous over time. Furthermore, the developed Fourier-MIONet makes it possible to solve the long-time evolution of geological carbon sequestration in a large-scale three-dimensional space accurately and efficiently.

97 MATHEMATICS AND COMPUTING↗

Enabling Seamless Transitions from Experimental to Production HPC for Interactive Workflows

The evolving landscape of scientific computing requires seamless transitions from experimental to production HPC environments for interactive workflows. This paper presents a structured transition pathway developed at OLCF that bridges the gap between development testbeds and production systems. We address both technological and policy challenges, introducing frameworks for data streaming architectures, secure service interfaces, and adaptive resource scheduling for time-sensitive workloads and improved HPC interactivity. Our approach transforms traditional batch-oriented HPC into a more dynamic ecosystem capable of supporting modern scientific workflows that require near real-time data analysis, experimental steering, and cross-facility integration.

Etz, Brian [ORNL] (ORCID:0000000208554863)↗

eCounter: Inline Per-IP Network Monitoring at Millisecond Resolution via eBPF

Scientific data acquisition (SciDAQ) systems are shifting from archive-based workflows to streaming paradigms, where real-time, fine-grained network monitoring becomes essential. While P4-enabled devices offer per-packet in-band observability, they require specialized switches and routers. Host-side tools like Prometheus exporters lack sufficient temporal granularity. To bridge this gap, we present eCounter, a lightweight, hardware-agnostic, inline telemetry agent built on extended Berkeley Packet Filter (eBPF). eCounter captures per-interface ingress and egress traffic, categorized by IP address and protocol, at millisecond to sub-millisecond resolution. In a 100 Gbps environment, it continuously exports up to 3,257 time-series bins per second with only 4% CPU utilization at a 35¿KiB/s data rate. We evaluate eCounter across diverse NIC MTU settings, hook types, CPU architectures and operating systems, and observed negligible impact on concurrent high-throughput streaming applications. Complexity analysis confirms that it can be readily scaled to distributed SciDAQ deployments.

Mei, Xinxin [Computational Sciences and Technology↗

Demonstration of Cross-Resonance Gates with Resonator-Assisted ZZ Cancellation

We present the characterization of a CNOT gate realized by combining cross-resonance interaction with resonator-assisted ZZ cancellation in fixed-frequency transmons on a Rigetti–SQMS co-developed quantum processor. Extending earlier work on dynamical ZZ cancellation via off-resonant resonator drives [1], we demonstrate a direct CNOT gate implementation achieved through two microwave drives on the transmons that generate a CX rotation in the |10⟩−|11⟩ subspace while selectively darkening the |00⟩−|01⟩ transition. This tunable-coupler-free approach enables high-fidelity gates and enhances the scalability of superconducting quantum architectures. [1] Z. Huang et al., Phys. Rev. Applied 22, 034007 (2024)

Heidler, Paul [Fermilab]↗

Pulsed Infrared Thermography Nondestructive Imaging of SiC-SiC f Composite Cladding Architectures; Understanding the Performance of SiC-SiC f Composite Cladding Architectures with Cr Coating in Normal Operating and Accident Conditions in LWRs and Advanced Reactors

SiC-SiC f composites, consisting of silicon carbide fibers embedded in a silicon carbide matrix, are advanced materials with high thermal conductivity, temperature stability, and resistance to radiation damage. Traditional methods for quality control of fabricated SiC-SiC f composites involve nondestructive evaluation (NDE) with X-ray computed tomography (XCT). However, XCT imaging of typical SiC-SiC f structures for cladding applications can involve several hours. In this project, we investigate an alternative approach to NDE of SiC-SiC f composites that involves rapid (on the order of seconds) imaging with Pulsed infrared thermography (PIT). PIT images of a planar SiC-SiC f specimen show the structure of the surface monolith layer and internal SiC f structures. The capability of PIT imaging in visualizing SiC f structures is qualitatively confirmed by observing similarity in the PIT and X-ray transmission images of the same specimen. Computer vision analysis of defects in the PIT image of the monolith was performed with thresholding followed by topological structural analysis that computed geometric descriptors, including major/minor axes of fitted ellipses, area, perimeter, Feret diameter, circularity, roundness, and solidity.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗