Search NASASearch

SEARCH · Search NASA

Results for “Computer systems performance”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 127 records · Page 7

How to Build a Quantum Supercomputer: Scaling from Hundreds to Millions of Qubits

In the span of four decades, quantum computation has evolved from an intellectual curiosity to a potentially realizable technology. Today, small-scale demonstrations have become possible for quantum algorithmic primitives on hundreds of physical qubits and proof-of-principle error-correction on a single logical qubit. Nevertheless, despite significant progress and excitement, the path toward a full-stack scalable technology is largely unknown. There are significant outstanding quantum hardware, fabrication, software architecture, and algorithmic challenges that are either unresolved or overlooked. These issues could seriously undermine the arrival of utility-scale quantum computers for the foreseeable future. Here, we provide a comprehensive review of these scaling challenges. We show how the road to scaling could be paved by adopting existing semiconductor technology to build much higher-quality qubits, employing system engineering approaches, and performing distributed quantum computation within heterogeneous high-performance computing infrastructures. These opportunities for research and development could unlock certain promising applications, in particular, efficient quantum simulation/learning of quantum data generated by natural or engineered quantum systems. To estimate the true cost of such promises, we provide a detailed resource and sensitivity analysis for classically hard quantum chemistry calculations on surface-code error-corrected quantum computers given current, target, and desired hardware specifications based on superconducting qubits, accounting for a realistic distribution of errors. Furthermore, we argue that, to tackle industry-scale classical optimization and machine learning problems in a cost-effective manner, heterogeneous quantum-probabilistic computing with custom-designed accelerators should be considered as a complementary path toward scalability.

Mohseni, Masoud

Hydra: computer vision for data quality monitoring

Hydra, initially developed for Hall-D in 2019, is a system that utilizes computer vision to perform near real time data quality monitoring. Since then, it has been deployed across all experimental halls at Jefferson Lab, with the CLAS12 collaboration in Hall-B being the first outside of GlueX to fully utilize Hydra. The system comprises back end processes that manage the models, their inferences, and the data flow. Finally, the front-end components, accessible via web pages, allow detector experts and shift crews to view and interact with the system.

47 OTHER INSTRUMENTATION

COLUMBUS─An Efficient and General Program Package for Ground and Excited State Computations Including Spin–Orbit Couplings and Dynamics

The COLUMBUS program system provides the tools for performing high-level multireference (MR) computations, including the multireference configuration interaction (MRCI) method and its multireference averaged quadratic coupled cluster (MR-AQCC) extension, allowing computations on a wide range of fascinating atomic and molecular systems, including the treatment of open-shells and complicated excited state phenomena. The inclusion of spin−orbit coupling (SOC) directly within the MRCI step enables the description of systems containing heavy elements, such as lanthanides and actinides, whose properties are strongly influenced by SOC. Analytic energy gradients and nonadiabatic couplings at the correlated MRCI level provide the foundation for a variety of dynamics studies, giving insight into ultrafast photochemistry. New and ongoing method developments in COLUMBUS include the computation of spin densities, improved descriptions of ionic states, enhancements to the AQCC method, and the porting of COLUMBUS to graphical processing units (GPUs). New external interfaces enable an enhanced description of electronic resonances and molecules in strong laser fields. This work highlights these new developments while providing a detailed account of the diverse applications of COLUMBUS in recent years.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Nonconvex Robust Optimization for the Design and Operation of Advanced Energy Systems Using PyROS

This work discusses recent advances of the two-stage robust optimization (RO) solver PyROS and applications to advanced energy systems optimization. To demonstrate the computational performance and reliability of PyROS, a study on a monoethanolamine (MEA)-based CO2 absorption flowsheet is presented. (Near-)robust feasible designs for CO2 absorption flowsheet at high carbon capture are obtained with the PyROS solver. The results demonstrate that the PyROS solver, including recent extensions to multi-stage RO settings, provides a reliable avenue to optimize the design and operation of advanced energy systems subject to various sources of parametric uncertainty.

Sherman, Jason

Diaspora: Resilience-Enabling Services for Real-Time Distributed Workflows

The need for real-time processing to enable automated decision making and experimental steering has driven a shift from high-performance computing workflows on a centralized system to a distributed approach that integrates remote data sources, edge devices, and diverse compute facilities. Under this paradigm, data can be processed close to the source where it is generated, thus reducing latency and bandwidth usage. System resilience is thus a key challenge, requiring distributed workflows to survive component failures and to meet stringent quality-of-service requirements, which results in the need to mitigate anomalies such as congestion and low availability of resources. To address these challenges, we propose Diaspora, a unified resilience framework that is inspired by event-driven communication patterns used in public clouds. Specifically, we propose an event fabric that extends across sites, facilities, and computations to provide timely, reliable, and accurate information about data, application, and resource status. On top of the event fabric, we build resilience-enabling services that combine QoS-aware data streaming, resilient data views, resilient compute and data resources, and anomaly detection and prediction, all of which collectively enhance workflow resilience for these scientific cases.

Rao, Nageswara

Toward integrating high-fidelity CFD approaches in the thermal-hydraulic analysis of turbulent dry cask systems

Nuclear power plants have been supplying resilient and reliable electricity for decades, contributing to energy independence of the U.S.. However, nuclear waste management remains one of the most significant challenges in the industry. The safety of dry cask storage systems relies heavily on their thermal-hydraulic performance. Computational Fluid Dynamics (CFD) simulations are often used to demonstrate this performance and ensure that the system design meets safety standards. This study presents reduced numerical models for various types of dry cask systems. These numerical models can produce efficient and fast results based on the employed modeling strategies. Additionally, the study uses a novel approach to high-fidelity simulations to evaluate modeling assumptions in dry cask modeling. Large Eddy Simulations (LES) are used for this purpose, particularly in regions where fluid velocity is relatively high and the turbulence characteristics become important. Furthermore, the results of these high-fidelity simulations will enhance the interpretation of outcomes produced from a lower-fidelity CFD model.

CFD

PaRSEC: Scalability, flexibility, and hybrid architecture support for task-based applications in ECP

This paper highlights the most significant enhancements made to PaRSEC, a scalable task-based runtime system designed for hybrid machines, during the Exascale Computing Project (ECP). The enhancements focus on expanding the capabilities of PaRSEC to address the evolving landscape of parallel computing. Notable achievements include the integration of support for three major types of accelerators (NVIDIA, AMD, and Intel GPUs), the refinement and increased flexibility of the communication subsystem, and the introduction of new programming interfaces tailored for irregular applications. Additionally, the project resulted in the development of powerful debugging and performance analysis tools aimed at assisting users in understanding and optimizing their applications. We present a comprehensive demonstration of these advancements through a series of benchmarks and applications within ECP and beyond, thereby showcasing the enhanced capabilities of PaRSEC across the diverse architectures within the ECP, providing valuable insights into the runtime system’s adaptability and performance across varied computing environments.

Bouteiller, Aurelien

Automation for Grid Interconnected Laboratory Emulation

As computational capabilities improve, digital twins are becoming vital for evaluating equipment realistically in laboratories. This paper outlines a digital twin architecture for the power grid, employing electromagnetic transient (EMT) simulation alongside real-time simulation of power hardware and hierarchical control systems. EMT simulation occurs on a high-performance computing server for scalability. Additionally, the paper describes a workflow and real-time data streaming software facilitating connectivity among EMT simulation, hierarchical control systems, and power hardware. This software enables automated equipment connectivity in the laboratory for realistic evaluations, aiding in identifying necessary upgrades for both equipment control systems and the power grid.

Marthi, Phani Ratna Vanamali [ORNL] (ORCID:0000000

Transforming Energy Through Computational Excellence: NREL's Computational Science Center

Computational methods underpin advancing the science and engineering of energy efficiency, sustainable transportation, renewable power technologies, and developing a knowledge base to optimize energy systems. NREL's Computational Science Center (CSC) proudly focuses on providing the service of computing, advancing the science of computing, and enabling NREL's clean energy mission.

applied mathematics

Bridging paradigms: Designing for HPC-Quantum convergence

Here, this paper presents a comprehensive software stack architecture for integrating quantum computing (QC) capabilities with High-Performance Computing (HPC) environments. While quantum computers show promise as specialized accelerators for scientific computing, their effective integration with classical HPC systems presents significant technical challenges. We propose a hardware-agnostic software framework that supports both current noisy intermediate-scale quantum devices and future fault-tolerant quantum computers, while maintaining compatibility with existing HPC workflows. The architecture includes a quantum gateway interface, standardized APIs for resource management, and robust scheduling mechanisms to handle both simultaneous and interleaved quantum–classical workloads. Key innovations include: (1) a unified resource management system that efficiently coordinates quantum and classical resources, (2) a flexible quantum programming interface that abstracts hardware-specific details, (3) A Quantum Platform Manager API that simplifies the integration of various quantum hardware systems, and (4) a comprehensive tool chain for quantum circuit optimization and execution. We demonstrate our architecture through implementation of quantum–classical algorithms, including the variational quantum linear solver, showcasing the framework’s ability to handle complex hybrid workflows while maximizing resource utilization. This work provides a foundational blueprint for integrating QC capabilities into existing HPC infrastructures, addressing critical challenges in resource management, job scheduling, and efficient data movement between classical and quantum resources.

97 MATHEMATICS AND COMPUTING

An Integrated Framework for Memory-Centric Analysis: From Trace Collection to Co-Design

The memory wall phenomenon—where advances in processor performance significantly outpace those in memory subsystems—poses a fundamental challenge for contemporary computing systems. In memory-bound applications, memory subsystem behavior dominates performance, yet existing analysis approaches present significant limitations: detailed microarchitectural simulators require days to weeks to simulate modest workloads; hardware performance counters provide only aggregate statistics that obscure temporal and spatial access patterns; and scaled simulation approaches face challenges in capturing certain behaviors that emerge at larger scales. These limitations reflect a processor-centric design philosophy increasingly misaligned with memory-bound workloads where detailed understanding of memory access patterns, cache hierarchy interactions, and contention is critical for effective optimization. This paper presents an integrated framework for memory-centric analysis that enables effective hardware-software co-design. We describe practical trace collection techniques, including hardware-assisted processor tracing with minimal overhead and portable software-based instrumentation with statistical sampling. We present multi-perspective analysis methods that examine memory behavior from temporal, sequential, spatial, and relational viewpoints, revealing distinct optimization opportunities invisible in aggregate metrics. We detail an architectural modeling framework that uses sampled traces with temporal interpolation and confidence-based filtering to evaluate cache and memory configurations. Evaluation on representative benchmarks demonstrates that this framework achieves practical accuracy (L2 cache errors of 2.64\%, confidence-filtered L3 errors of 9.92\%, bandwidth errors of 7.33\%) while providing substantial speedup (26.8×) over cycle-accurate simulation, enabling rapid design space exploration. We demonstrate how this integrated framework enables systematic identification of both hardware optimizations (memory controller tuning, bank partitioning, NUMA configuration) and software optimizations (data layout restructuring, prefetching strategies, memory-aware scheduling). Through this comprehensive treatment of the memory-centric analysis pipeline—from trace collection through architectural modeling to co-design application—we provide researchers and practitioners with practical techniques for addressing memory bottlenecks in contemporary computing systems.

Gajaria, Dhruv Mayur

Frontiers in Scientific Workflows: Pervasive Integration With High-Performance Computing

Herein we address the increasing complexity of scientific workflows in the context of high-performance computing (HPC) and their associated need for robust, adaptable, and flexible computational support systems. We explore five key trends as well as future challenges and opportunities for scientific workflows and HPC technologies.

97 MATHEMATICS AND COMPUTING

An improved guess for the variational calculation of charge-transfer excitations in large systems

Ab initio quantum-chemical methods that perform well for computing the electronic ground state are not straightforwardly transferable to electronically excited states, particularly in large molecular systems. Wave function theory offers high accuracy, but is often prohibitively expensive. Methods based on time-dependent density functional theory (TD-DFT) are crucially sensitive to the chosen exchange-correlation functional (XCF) parameterization, and system-specific tuning protocols were therefore proposed to address the method's robustness. Methods based on the variational relaxation of the excited-state electron density showcased promising results for the calculation of charge-transfer excitations, but the complex shape of the electronic hypersurface makes convergence to a specific excited state much more difficult than for the ground state when standard variational techniques are applied. We address the latter aspect by providing suitable initial guesses, which we obtain by two separate constrained algorithms. Combined with the squared-gradient minimization algorithm for all-electrons relaxation in a freeze-and-release scheme (FRZ-SGM), we demonstrate that orbital-optimized density functional theory (OO-DFT) calculations can reliably converge to the charge-transfer states of interest even for large molecular systems. We test the FRZ-SGM method on a phenothiazine-anthraquinone CT excitation in a supramolecular Pd(II) coordination cage complex as a function of the cage conformation. This compound has been studied experimentally prior to our work. We compare this freeze-and-release scheme to two XCF reparameterizations, which were recently proposed as low-cost TD-DFT-based alternatives to variational methods. Two dye-semiconductor complexes, which were previously investigated in the context of photovoltaic applications, serve as a second example to investigate the convergence and stability of the FRZ-SGM approach. Our results demonstrate that FRZ-SGM provides reliable convergence for charge-transfer excited states and avoids variational collapse to lower-lying electronic states, whereas time-dependent DFT calculations with an adequate tuning procedure for the range-separation parameter provide a computationally efficient initial estimate of the corresponding energies, with a computational cost comparable to that of configuration-interaction singles (CIS) calculations.

Bogo, Nicola

A Hands-On Curriculum for Training in HPC Cluster Deployment and Management

This paper presents the design, methodology, and outcomes of the High-Performance Computing Technologies (HPCT) course, a hands-on training program focused on the system-side of HPC cluster deployment and administration. Delivered as part of the Master in High Performance Computing (MHPC) program, the course introduces students to key concepts in cluster configuration, including networking, software stack provisioning, job scheduling, and monitoring. Initially taught in person, the course was transitioned to an online format during the COVID-19 pandemic. This shift led to the development of openly available instructional material and a flipped-classroom approach that continues to support both in-person and hybrid delivery. All course materials are publicly available at www.hpc.temple.edu/mhpc/hpc-technology/index.html. By documenting the structure, infrastructure, and evolution of HPCT, this paper offers a model for accessible HPC system training that supports workforce development in computational science.

Posada Correa, Fernando [ORNL] (ORCID:000000022565

Fabrication and Testing of Solid-Solution Strengthened Corrosion Resistant Alloys For Service in Molten Fluoride Environments

The demand for higher system thermal efficiencies requires the operation of power generation cycles and heat conversion systems at progressively higher temperatures. As the system operating temperature increases, existing materials may not provide adequate mechanical properties or environmental compatibility or both. There is an increasing commercial interest in the development and deployment of liquid-fueled Molten Salt Reactors (MSRs). Hastelloy®N, the highest performing candidate MSR structural alloy, is not capable of operations at temperatures above 700°C, thus limiting the performance of these systems. Using an Integrated Computational Materials Engineering (ICME)-approach and small laboratory scale heats, ORNL developed a class of patented alloys covered by U.S. Patent 9, 435, 011 B2, “Creep-resistant, Cobalt-free alloys for high temperature, liquid-salt heat exchanger systems,” similar to Hastelloy®N in that they are primarily solid solution strengthened. In contrast to precipitation strengthened alloys, the microstructure of solid solution alloys and hence the high temperature mechanical properties are stable for extended periods of time allowing long reactor operating life. The new alloys have shown to possess good resistance to liquid fluorides at temperatures up to 850°C and have significantly improved creep properties when compared to Hastelloy®N. The purpose of the CRADA project was for ORNL to collaborate with Haynes International- a materials producer, MetalTek International- a foundry, and Kairos Power – an advanced reactor developer – to scale-up selected alloys, evaluate their properties, and identify one solid solution strengthened alloy that can meet the property requirements for the reactor being developed by Kairos Power and other similar liquid fluoride-salt cooled reactors. As part of the project, eight alloys were down-selected and fabricated in larger industrial scale heats by Haynes International. Resistance to molten salt was evaluated in flowing FLiNaK and FLiBe by Kairos Power using their Rotating Cage Loop (RCL) system. Accounting for iron deposition during these tests, the new alloys displayed very low net mass change showing excellent corrosion performance in molten salt. Creep properties evaluated at ORNL were found to be better than that of Hastelloy®N and 316 stainless steel. Long-term stabilities of the alloys evaluated by Haynes International showed that these alloys have excellent thermal stability in the temperature range 704.4-815.6°C, with the change in strength and ductility being less than 10-15% after a 4000-hour exposure at 815.6°C. Autogenously Gas Tungsten Arc Welding (GTAW) welded samples showed less than 10% change in yield strength / ultimate tensile strength / total elongation compared to the basemetal, indicating that the alloys have excellent weldability. Three parts were successfully investment-cast using one alloy with very little voiding showing feasibility of fabricating parts using the casting process. This project enabled extensive interaction between the material producer Haynes International, casting supplier MetalTek, and reactor developer Kairos Power. This facilitated testing of materials and components produced using the newly developed alloys by the end-user. This allowed the generation of critical dataset required for down-selection of a few promising alloys for further development. This data is also currently being shared with other reactor designers for them to evaluate the suitability of this alloy for their reactor design. The availability of this alloy will ultimately enable the design and development and deployment of MSRs with increased temperature of operation and thus, improved efficiencies.

99 GENERAL AND MISCELLANEOUS

Fabrication and Testing of Solid-Solution Strengthened Corrosion Resistant Alloys For Service in Molten Fluoride Environments

The demand for higher system thermal efficiencies requires the operation of power generation cycles and heat conversion systems at progressively higher temperatures. As the system operating temperature increases, existing materials may not provide adequate mechanical properties or environmental compatibility or both. There is an increasing commercial interest in the development and deployment of liquid-fueled Molten Salt Reactors (MSRs). Hastelloy®N, the highest performing candidate MSR structural alloy, is not capable of operations at temperatures above 700°C, thus limiting the performance of these systems. Using an Integrated Computational Materials Engineering (ICME)-approach and small laboratory scale heats, ORNL developed a class of patented alloys covered by U.S. Patent 9,435,011 B2, “Creep-resistant, Cobalt-free alloys for high temperature, liquid-salt heat exchanger systems,” similar to Hastelloy®N in that they are primarily solid solution strengthened. In contrast to precipitation strengthened alloys, the microstructure of solid solution alloys and hence the high temperature mechanical properties are stable for extended periods of time allowing long reactor operating life. The new alloys have shown to possess good resistance to liquid fluorides at temperatures up to 850°C and have significantly improved creep properties when compared to Hastelloy®N. The purpose of the CRADA project was for ORNL to collaborate with Haynes International- a materials producer, MetalTek International- a foundry, and Kairos Power – an advanced reactor developer – to scale-up selected alloys, evaluate their properties, and identify one solid solution strengthened alloy that can meet the property requirements for the reactor being developed by Kairos Power and other similar liquid fluoride-salt cooled reactors. As part of the project, eight alloys were down-selected and fabricated in larger industrial scale heats by Haynes International. Resistance to molten salt was evaluated in flowing FLiNaK and FLiBe by Kairos Power using their Rotating Cage Loop (RCL) system. Accounting for iron deposition during these tests, the new alloys displayed very low net mass change showing excellent corrosion performance in molten salt. Creep properties evaluated at ORNL were found to be better than that of Hastelloy®N and 316 stainless steel. Long-term stabilities of the alloys evaluated by Haynes International showed that these alloys have excellent thermal stability in the temperature range 704.4-815.6°C, with the change in strength and ductility being less than 10-15% after a 4000-hour exposure at 815.6°C. Autogenously Gas Tungsten Arc Welding (GTAW) welded samples showed less than 10% change in yield strength / ultimate tensile strength / total elongation compared to the basemetal, indicating that the alloys have excellent weldability. Three parts were successfully investment-cast using one alloy with very little voiding showing feasibility of fabricating parts using the casting process. This project enabled extensive interaction between the material producer Haynes International, casting supplier MetalTek, and reactor developer Kairos Power. This facilitated testing of materials and components produced using the newly developed alloys by the end-user. This allowed the generation of critical dataset required for down-selection of a few promising alloys for further development. This data is also currently being shared with other reactor designers for them to evaluate the suitability of this alloy for their reactor design. The availability of this alloy will ultimately enable the design and development and deployment of MSRs with increased temperature of operation and thus, improved efficiencies.

36 MATERIALS SCIENCE

Resilience of the Electric Grid Through Trustable IoT-Coordinated Assets

The electricity grid has evolved from a physical system to a cyberphysical system with digital devices that perform measurement, control, communication, computation, and actuation. The increased penetration of distributed energy resources (DERs) including renewable generation, flexible loads, and storage provides extraordinary opportunities for improvements in efficiency and sustainability. However, they can introduce new vulnerabilities in the form of cyberattacks, which can cause significant challenges in ensuring grid resilience. We propose a framework in this paper for achieving grid resilience through suitably coordinated assets including a network of Internet of Things devices. A local electricity market is proposed to identify trustable assets and carry out this coordination. Situational Awareness (SA) of locally available DERs with the ability to inject power or reduce consumption is enabled by the market, together with a monitoring procedure for their trustability and commitment. With this SA, we show that a variety of cyberattacks can be mitigated using local trustable resources without stressing the bulk grid. Multiple demonstrations are carried out using a high-fidelity cosimulation platform, real-time hardware-in-the-loop validation, and a utility-friendly simulator.

distributed energy resources

Regen: An object layout regenerator on large-scale production HPC systems

This article proposes an object layout regenerator called Regen which regenerates and removes the object layout dynamically to improve the read performance of applications. Regen first detects frequent access patterns from the I/O requests of the applications. Second, Regen reorganizes the objects and regenerates or preallocates new object layouts according to the identified access patterns. Finally, Regen removes or reuses the obsolete or regenerated object layouts as necessary. As a result, Regen accelerates access to objects by providing a flexible object layout. We implement Regen as a framework on top of Proactive Data Container (PDC) and evaluate it on Cori supercomputer, a production-scale HPC system, by using realistic HPC I/O benchmarks. The experimental results show that Regen improves the I/O performance by up to 16.92 × compared with an existing system.

Distributed file system