Search NASA⌕ Search

SEARCH · Search NASA

Results for “high performance”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Electric Drive Technologies Consortium (EDTC)/ Cost competitive, high-Performance, highly Reliable (CPR) Power Devices on 4H-SiC (Final Report)

4H-Silicon carbide (4H-SiC) is a wide bandgap semiconductor that offers superior material properties over silicon, including higher critical electric field, thermal conductivity, and electron saturation velocity. These advantages make 4H-SiC highly attractive for high-voltage, high-efficiency power electronics. However, realizing the full potential of SiC requires device technologies that are not only high-performing but also manufacturable and reliable under real-world operating conditions. This report summarizes the outcomes of a five-year R&D effort funded by the U.S. Department of Energy (DOE) under the Electric Drive Technologies Consortium (EDTC), focused on developing cost-competitive, high-performance, and highly reliable (CPR) power devices on 4H-SiC substrates. The program targeted scalable and manufacturable 1.2 kV-class SiC MOSFETs optimized for next-generation electric vehicles, renewable energy systems, and industrial power conversion. The project delivered transformative advancements in SiC power device performance and ruggedness. Particularly, Specific on-resistance (R on,sp ) was reduced by up to 37%, from ~4.0 m$\Omega \cdot$cm 2 in earlier designs to an industry-leading 2.40 m$\Omega \cdot$cm 2 , driven by optimized doping, refined JFET widths, and layout engineering. Breakdown voltages (BV) exceeded 1600 V, marking improvement over legacy baselines, and demonstrating the robustness of newly implemented junction profiles and edge terminations. Short-circuit withstand time (SCWT) saw a remarkable 4$\times$ increase, from ~2 $\mu$s to over 8 $\mu$s, achieved through the successful deployment of deep P-well structures (~1.8–2.0 $\mu$m) via channeling implantation. This innovative process breakthrough enabled precise junction formation without MeV-class implantation tools, reduced leakage under high field stress, and allowed even the shortest-channel devices (down to 0.3 $\mu$m) to achieve both high BV and excellent ruggedness—breaking the traditional trade-off between conduction efficiency and blocking capability. Several novel architectures pushed the performance envelope further. JBSFETs—featuring embedded Schottky portions—eliminated bipolar degradation and drastically reduced third-quadrant leakage, while Ladder MOSFETs introduced a clever orthogonal conduction path that achieved a 15.4% reduction in R on,sp over standard linear designs. Switching performance reached new benchmarks: short-channel devices showed a 31% reduction in total switching energy compared to 0.5 $\mu$m counterparts, while maintaining manageable gate drive requirements. Layout-optimized structures not only improved transconductance but also accelerated switching transitions, pointing to real-world benefits in converter-level efficiency. The devices also passed rigorous reliability validation. Stress-tested across TDDB, HTGB, HTRB, HVP, and burn-in, the devices screened under 30 V/10 hr and 43 V/1 s protocols consistently exhibited tighter lifetime distributions and long-term oxide robustness. These screening techniques proved effective in identifying latent defects and ensuring deployment-grade reliability. Meanwhile, advanced 3D TCAD simulations revealed and resolved electric field hotspots—particularly in HEXFET corners—where fields exceeding 4.8 MV/cm were mitigated through geometry-aware layout corrections. Overall, the results of this project demonstrate a manufacturable and scalable SiC power device platform that addresses key DOE performance targets for efficient, robust, and reliable 1.2kV 4H-SiC Power Devices. The developed technologies represent a meaningful step forward in the commercial readiness of high-voltage SiC solutions and provide a strong foundation for continued advancement in wide bandgap power electronics.

42 ENGINEERING↗

Fast ion studies in the extended high-performance high β P plasma on EAST

Comprehending and optimizing fast ion behaviors is critical for the enhancement of performance in Experimental Advanced Superconducting Tokamak (EAST). This study explores the potential benefits of several factors that can improve the fast ion confinement. First, experiments show the change in the direction of the NBI2 from counter-I p to co-I p leads to a significant reduction in fast ion losses. TRANSP/NUBEAM simulation and tomography results based on fast-ion D-alpha measurements reveal that after the neutral beam injection (NBI) upgrade, the beam ion prompt loss is reduced by approximately 50%. Second, the upgraded ion cyclotron resonant frequency (ICRF) antenna at the N-port features twice the coupling resistance of the original antennas at EAST. This improved ICRF power coupling has enhanced the synergistic heating effect of NBI + ICRF, where the ICRF wave field accelerates beam ions at the harmonics. Experiments demonstrate that NBI + ICRF synergistic not only enhances plasma neutron yield and β P , but also accelerates beam ions to hundreds of keV. Further, the electron density and the neutral beam voltage have been optimized to reduce the fast ion slowing-down time and beam ion losses. Experimental and simulation results indicate that increasing the electron density reduces beam ion losses and enhances the bootstrap current fraction. While higher beam voltage results in a slight decrease in beam power absorption, it can increase the fraction of bootstrap current. With the understanding of these optimization of fast ion confinement, experiments have demonstrated fully non-inductive operation at high density (n e /n G ∼ 0.67, β P ∼ 3.1, β N ∼ 2.1, H 98,y2 ∼ 1.2) even without the support of co-I p beam NBI2. This investigation presents a potential regime to enhance fast ion confinement and extend performance in the high β P plasma for future experiments.

EAST tokamak↗

High Performance, High Fidelity: A GPU‐Accelerated Doubly‐Periodic Configuration of the Simple Cloud‐Resolving E3SM Atmosphere Model Version 1 (DP‐SCREAMv1)

The development of the Simplified Cloud Resolving Energy Exascale Earth System Atmosphere Model (SCREAMv1) enables global storm-resolving simulations on modern GPU-based supercomputers. However, the high computational cost of SCREAMv1 limits its routine use for process-level studies, creating a need for efficient proxy configurations. This study addresses this gap by introducing DP-SCREAMv1, a doubly periodic cloud-resolving model designed to be fully consistent with SCREAMv1 while enabling high-resolution, long-duration simulations at significantly reduced computational expense by simulating a limited doubly periodic domain rather than the entire globe. Built on a C++/Kokkos architecture, DP-SCREAMv1 achieves exceptional performance scalability on GPU systems and includes a rich library of cases for validation and scientific exploration. In this work, we demonstrate short wall-clock times at SCREAMv1's default resolution and show that DP-SCREAMv1 supports routine execution of large-domain, high-resolution experiments that were previously challenging in practice. Furthermore, we show that DP-SCREAMv1 enables routine execution of “Giga-LES” style simulations and facilitates large-domain, high-resolution simulations that were recently considered burdensome to perform. These results document an efficient, fully consistent process-level configuration for SCREAMv1 (DP-SCREAMv1) and illustrate its use for long-duration and large-domain experiments at cloud-resolving to eddy-permitting resolution.

Environmental sciences↗

Accurately constituting robust interfaces for high-performance high-energy lithium metal batteries

High-energy lithium metal batteries (LMBs) have received ever-increasing interest. Among them, coupling lithium metal (Li) with nickel-rich material, LiNi x Mn y Co z O 2 (NMCs, x ≥ 0.6, x + y + z = 1), is promising because Li anodes enable an extremely high capacity (∼3860 mA h g −1 ) and the lowest redox potential (−3.04 V vs. standard hydrogen electrode), while NMCs can achieve a much higher capacity of ∼200 mA h g −1 and lower cost than those of LiCoO 2 . However, the resultant Li‖NMC cells have been hindered from commercialization due to a series of challenges related to the interface stability of both Li anodes and NMC cathodes. Specifically, Li anodes suffer from Li dendritic growth and the formation of solid electrolyte interphase (SEI), while NMC cathodes suffer from the formation of cathode electrolyte interphase (CEI) and other interface-related issues, including transition metal dissolution, oxygen release, cracking, and so on. To tackle these issues, recently, two sister techniques, atomic and molecular layer deposition (ALD and MLD), have emerged and exhibit tremendous capabilities to accurately constitute robust interfaces to achieve high-performance Li‖NMC LMBs. They can uniquely develop uniform and conformal films as surface coatings of LMBs in a precisely controllable mode at the atomic/molecular level, while proceeding with film deposition at low temperatures (e.g., ≤250 °C). In this Feature Article, we review the latest research progress in developing novel surface coatings via ALD and MLD for Li‖NMC LMBs and discuss outcomes for pursuing high performance.

25 ENERGY STORAGE↗

Towards Autonomous Experiments by Connecting High Performance Microscopy with High Performance Computing

The digitization of controls, data, and analysis in microscopy is bringing the idea of autonomous microscopes closer to reality than ever before. Automated transmission electron microscopy (TEM) is already fairly routine for some experiments the only require simple repetitive tasks such as imaging biological macromolecules for single particle cryoEM [1], tilt series for electron tomography [2], and movies for crystallography [3]. The vast majority of TEM experiments are conducted completely by human operators who choose the regions of interest, optimize experimental parameters, and make decisions about data quality visually during an experiment. The field is still a long way from having completely autonomous TEMs that can adapt to sample difficulties and tune experimental parameters based on data quality and desired experimental outcomes. Part of the issue is the lack of capability for feeding information learned from on-line, live data analysis back into the on-going experiment [4]. Furthermore, this presentation will discuss current capabilities for large scale data reduction and analysis using high performance computing (i.e. supercomputing) and progress towards developing a true feed-back loop that places data analysis and theory in the experimental loop.

97 MATHEMATICS AND COMPUTING↗

Experimental Validation of a Kinetic Ballooning Mode in High-Performance High-Bootstrap Current Fraction Fusion Plasmas

We report the observation of a set of coherent high frequency electromagnetic fluctuations that leads to a turbulence induced self-regulating phenomenon in the DIII-D high bootstrap current fraction plasma. The fluctuations have frequency of 130~220kHz, the poloidal wave length and phase velocity are 16~30 m -1 and ~30 km/s, respectively in the outboard midplane with the estimated toroidal mode number n~5- 9. The fluctuations are located in the internal transport barrier (ITB) region at large radius and are experimentally validated to be kinetic ballooning modes (KBM). Furthermore, quasilinear estimation predicts the KBM to be able to drive experimental particle flux and non-negligible thermal flux, suggesting its significant role in regulating the ITB saturation.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Interpretable Models for Workflow Differentiation in High-Performance Scientific Networks

Scientific workflows in high-performance networks spawn hundreds of interdependent flows that must be managed collectively—yet existing network classifiers treat each flow in isolation, leading to fragmented QoS decisions and missed interflow patterns. We present a novel traffic classification solution that operates at the workflow level, distinguishing entire filetransfer operations from streaming analytics by capturing how concurrent flows interact and burst together. We introduce a workflow identification window (WIW) that ingests raw packet headers from parallel flows into unified tensors, preserving the spatial-temporal patterns that differentiate scientific workflows. This approach achieves 98.7% accuracy using CNN, LSTM, and hybrid architectures, while maintaining 84% accuracy on production traffic collected a week later—demonstrating robustness to temporal drift. By integrating SHAP and GradCAM explainability, we reveal that early-packet timing patterns and cross-flow correlations drive classification decisions, providing operators with interpretable insights. Our system enables coherent workflow-level QoS enforcement and dynamic bandwidth allocation in scientific networks, eliminating manual per-flow configuration while maintaining classification latency at millisecond level.

Giannakou, Anna [LBL, Berkeley]↗

Frontiers in Scientific Workflows: Pervasive Integration With High-Performance Computing

Herein we address the increasing complexity of scientific workflows in the context of high-performance computing (HPC) and their associated need for robust, adaptable, and flexible computational support systems. We explore five key trends as well as future challenges and opportunities for scientific workflows and HPC technologies.

97 MATHEMATICS AND COMPUTING↗

Toward a persistent event-streaming system for high-performance computing applications

High-performance computing (HPC) applications have traditionally relied on parallel file systems and file transfer services to manage data movement and storage. Alternative approaches have been proposed that use direct communications between application components, trading persistence and fault tolerance for speed. Event-driven architectures, as popularized in enterprise contexts, present a compelling middle ground, avoiding the performance cost and API constraints of parallel file systems while retaining persistence and offering impedance matching between application components. However, adapting streaming frameworks to HPC workloads requires addressing challenges unique to HPC systems. This paper investigates the potential for a streaming framework designed for HPC infrastructures and use cases. We introduce Mofka, a persistent event-streaming framework designed specifically for HPC environments. Mofka combines the capabilities of a traditional streaming service with optimizations tailored to the HPC context, such as support for massively multicore nodes, efficient scaling for large producer-consumer workflows, RDMA-enabled high-performance network communications, specialized network fabrics with multiple links per node, and efficient handling of large scientific data payloads. Built using the Mochi suite of HPC data service components, Mofka provides a lightweight, modular, and high-performance solution for persistent streaming in HPC systems. We present the architecture of Mofka and evaluate its performance against Kafka and Redpanda using benchmarks on diverse platforms, including Argonne's Polaris and Oak Ridge's Frontier supercomputers, showing up to 8× improvement in throughput in some scenarios. We then demonstrate its utility in several real-world applications: a tomographic reconstruction pipeline, a workflow for the discovery of metal-organic frameworks for carbon capture, and the instrumentation of Dask workflows for provenance tracking and performance analysis.

HPC↗

INL High-Performance and Sustainable Building Strategy

High-performance buildings are reliable, cost effective, and sustainable structures that minimize energy and water use, reduce solid waste and pollutant emissions, and limit the depletion of natural resources. High-performance buildings also provide a thermally and visually comfortable working environment that increases productivity for building occupants. As Idaho National Laboratory (INL) is the nation’s premier nuclear energy research laboratory, the physical infrastructure requires continual updating and repurposing to help accomplish that mission. INL’s infrastructure must incorporate high-performance sustainable design features to be fiscally responsible and reflect an image of innovation to the public and prospective employees. INL is a large consumer of energy with annual energy costs exceeding $16M. This High-Performance and Sustainable Building Strategy will help engineering and construction project teams design sustainable facilities, reduce life cycle operating costs, and support the INL net-zero plan while providing INL employees with a safe and healthy working environment. With these goals in mind, the recommendations described in this document are intended to form INL’s foundation for sustainable and high-performance building standards. This strategy incorporates the latest federal and Department of Energy (DOE) orders and directives, including DOE Order 436.1A, “Departmental Sustainability,” the DOE Sustainability Plan (SP), the INL Site Sustainability Plan (SSP), and Code of Federal Regulations (CFR). This document identifies the requirements of the “Guiding Principles for Sustainable Federal Buildings” (Guiding Principles) and briefly highlights the Leadership in Energy and Environmental Design (LEED) Gold certification. LEED Gold certification can be used to meet many of the requirements of the Guiding Principles.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

XaaS: Acceleration as a Service to Enable Productive High-Performance Cloud Computing

High-performance computing (HPC) and the cloud have evolved independently, specializing their innovations into performance or productivity. Acceleration as a Service (XaaS) is a recipe to empower both fields with a shared execution platform that provides transparent access to computing resources, regardless of the underlying cloud or HPC service provider. Bridging HPC and cloud advancements, XaaS presents a unified architecture built on performance-portable containers. Here, our converged model concentrates on low-overhead, high-performance communication and computing, targeting resource-intensive workloads from climate simulations to machine learning. XaaS lifts the restricted allocation model of Function as a Service (FaaS), allowing users to benefit from the flexibility and efficient resource utilization of serverless computing while supporting long-running and performance-sensitive workloads from HPC.

97 MATHEMATICS AND COMPUTING↗

High-Performance Lithium-Ion Batteries with High Stability Derived from Titanium-Oxide- and Sulfur-Loaded Carbon Spherogels

This study presents a novel approach to developing high-performance lithium-ion battery electrodes by loading titania-carbon hybrid spherogels with sulfur. The resulting hybrid materials combine high charge storage capacity, electrical conductivity, and core-shell morphology, enabling the development of next-generation battery electrodes. We obtained homogeneous carbon spheres caging crystalline titania particles and sulfur using a template-assisted sol-gel route and carefully treated the titania-loaded carbon spherogels with hydrogen sulfide. The carbon shells maintain their microporous hollow sphere morphology, allowing for efficient sulfur deposition while protecting the titania crystals. By adjusting the sulfur impregnation of the carbon sphere and varying the titania loading, we achieved excellent lithium storage properties by successfully cycling encapsulated sulfur in the sphere while benefiting from the lithiation of titania particles. Without adding a conductive component, the optimized material provided after 150 cycles at a specific current of 250 mA g -1 a specific capacity of 825 mAh g -1 with a Coulombic efficiency of 98%.

25 ENERGY STORAGE↗

Synthesis of Microscopic 3D Graphene for High‐Performance Supercapacitors with Ultra‐High Areal Capacitance

Abstract Despite graphene being considered an ideal supercapacitor electrode material, its use in commercial devices is limited because few methods exist to produce high‐quality graphene at a large scale and low cost. A simple method is reported to synthesize 3D graphene by graphenization of coal tar pitch with a K 2 CO 3 catalyst. This produces 3D graphenes with high specific surface areas up to 2113 m 2 g −1 and exceptional crystallinity (Raman I D / I G as low as ≈0.15). The material has an outstanding specific capacitance of 182.6 F g −1 at a current density of 1.0 A g −1 . This occurs at a mass loading of 30 mg cm −2 which is 3 times higher than commercial requirements, yielding an ultra‐high areal capacitance of 5.48 F cm −2 . The K 2 CO 3 is recycled and reused over 10 cycles with material quality and electrocapacitive performance of 3D graphene retained and verified after each cycle. The synthesis method and resulting electrocapacitive performance properties create new opportunities for using 3D graphene more broadly in practical supercapacitor devices.

36 MATERIALS SCIENCE↗

Matrix-Free High-Performance Saddle-Point Solvers for High-Order Problems in \(\boldsymbol{H}(\operatorname{\textbf{div}})\)

Here, this work describes the development of matrix-free GPU-accelerated solvers for high-order finite element problems in H(div). The solvers are applicable to grad-div and Darcy problems in saddle-point formulation, and have applications in radiation diffusion and porous media flow problems, among others. Using the interpolation–histopolation basis, efficient matrix-free preconditioners can be constructed for the (1, 1)-block and Schur complement of the block system. With these approximations, block-preconditioned MINRES converges in a number of iterations that is independent of the mesh size and polynomial degree. The approximate Schur complement takes the form of an M-matrix graph Laplacian and therefore can be well-preconditioned by highly scalable algebraic multigrid methods. High-performance GPU-accelerated algorithms for all components of the solution algorithm are developed, discussed, and benchmarked. Numerical results are presented on a number of challenging test cases, including the “crooked pipe” grad-div problem, the SPE10 reservoir modeling benchmark problem, and a nonlinear radiation diffusion test case.

97 MATHEMATICS AND COMPUTING↗

Queue wait time prediction in high performance computing (HPC) systems

High Performance Computing (HPC) systems are critical enablers for groundbreaking scientific research across various domains. Efficient resource allocation, facilitated by job scheduling, is paramount for maximizing the utilization of HPC systems. However, the variability in wait times for queued jobs poses challenges for users, necessitating accurate job wait time estimation. This paper explores the influence of job characteristics, including job size (the number of nodes requested and walltime), the queue to which the job is submitted and other resource requirements, on job wait times in leadership-class HPC systems. Focusing on the Theta Cray XC40 and Polaris machines at Argonne National Laboratory, the study evaluates the performance of different supervised learning algorithms in predicting job wait times. It also evaluates the impact of data preprocessing, including outlier detection, Principal Component Analysis (PCA), and feature selection, on the performance of wait time prediction models. The findings reveal insights into the relationship between job characteristics and wait times, offering a foundation for optimizing resource allocation and enhancing user experience. The methodologies and tools developed in this study are adaptable to other leadership-class HPC systems, providing a valuable contribution to the broader HPC community aiming to improve job scheduling efficiency and user satisfaction.

Okafor, Nwamaka↗

Data Center Immersion Cooling: A Case Study and Summary of High-Performance Computing Cooling Technologies

The future of High Performance Computing (HPC) is carved out for all high-performance facilities. As computer processing increases exponentially, especially with the demand for AI training, more speed will equate to more computing density and ultimately more power draw and resource demand. Sandia National Laboratories has not seen or known of anything that would lead us to think this might change in the next 5 to 10 years.

97 MATHEMATICS AND COMPUTING↗