Search NASA⌕ Search

SEARCH · Search NASA

Results for “Performance portability”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 163 records · Page 9

Spontaneous sodium ion storage behaviors of reduced graphene oxide anodes exceeding 100% Coulombic efficiency by modulated ion solvation

Rechargeable batteries are essential energy storage devices that power portable devices and electrical vehicles throughout the world. In general, it is thought that the electrochemical performance of rechargeable batteries is mostly determined by the electrodes within them and that the electrolyte plays a relatively passive role. However, ion transport and storage can be greatly influenced by the electrolyte solution structure, specifically, ion solvation within the bulk and ion desolvation across the electrode/electrolyte interfaces. Herein, we studied the role of the electrolyte as an active component of electrochemical energy storage devices. We found that with an appropriate electrolyte formulation, ion storage in disordered carbonaceous anode materials can occur spontaneously without externally supplied electrical energy. Reduced graphene oxide (RGO) in an ether-based electrolyte demonstrates ‘spontaneous' ion storage behaviors of adsorbing and inserting the solvated ions utilizing facilitated permeability and wettability of RGO, which results in Coulombic efficiency of ~145% due to additional charging capacity of ~180 mAh g -1 during electrochemical processes. The unexpected spontaneous ion storage behavior was extensively investigated using a combination of electrochemical analyses and diagnostics, advanced characterizations, and computational simulation. In conclusion, we believe the spontaneous ion storage behavior offers a new way to further improve the energy efficiency of practical rechargeable batteries.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Ku-band electron linac for battery-powered hand-portable 2-MeV X-ray generator

X-ray generators, producing radiation in MeV range, are a critical tool for radiography, non-destructive testing and security applications. Field operation of such source requires them to be hand-portable, autonomous and allow parameter adjustability. RF linear accelerators can serve as a flexible, reliable, and robust radiation generator alternative to dangerous radioisotopes and bulky betatrons that are currently used for field radiography if their size, weight, cost, and imaging performance are matched to these sources. Here, in this paper, we present the design and test results of a 2 MeV Ku-band electron linac for a hand-portable X-ray generator system for field radiography being developed by RadiaBeam. The dramatic scale of miniaturization and cost-reduction is achieved thanks to the implementation of innovative technologies such as air-cooled Ku-band air-traffic control magnetrons, a split accelerating structure fabrication technique, and solid-state Marx modulators. This paper presents the design of the first prototype of the accelerator, its operation from Li-Ion batteries, as well as high-power and beam measurements.

43 PARTICLE ACCELERATORS↗

Rapid Evaluation Framework for the CMIP7 Assessment Fast Track

As Earth system models (ESMs) grow in complexity and in volume of output data, there is an increasing need for rapid, comprehensive evaluation of their scientific performance. The upcoming Assessment Fast Track for the Seventh Phase of the Coupled Model Intercomparison Project (CMIP7) will require expeditious response for model analyses designed to inform and drive integrated Earth system assessments. To meet this challenge, the Rapid Evaluation Framework (REF), a community-driven platform for benchmarking and performance assessment of ESMs, was designed and developed. The initial implementation of the REF, constructed to meet the near-term needs of the CMIP7 Assessment Fast Track, builds upon four disparate community evaluation and benchmarking tools that are coupled together using the Coordinated Model Evaluation Capabilities (CMEC) framework. The REF runs within a containerized workflow for portability and reproducibility and is aimed at generating and organizing diagnostics covering a variety of model variables. The REF leverages well documented observational datasets to provide assessments of model fidelity across a collection of diagnostics. All diagnostics were identified and selected with community involvement and consultation. Operational integration with the Earth System Grid Federation (ESGF) will permit automated execution of the REF for selected diagnostics as soon as model output data are published on ESGF by the originating modeling centers. The REF is designed to be portable across a range of current computational platforms to facilitate use by modeling centers for assessing the evolution of model versions or gauging the relative performance of CMIP simulations before being published on ESGF. When integrated into production simulation workflows, results from the REF provide immediate quantitative feedback that allows model developers and scientists to quickly identify model biases and performance issues. After the REF is released to the community, its subsequent development and support will be prioritized by an international consortium of scientists and engineers, enabling a broader impact across Earth science disciplines. For instance, the REF will facilitate improvements to models and will enhance confidence in model projections through process-based selection of models based on their performance with respect to observations. Production of reproducible diagnostics and community-based assessments are key features of the REF. Furthermore, providing interoperability with existing evaluation packages assures that contributions from previous community efforts will be available for use in future model intercomparison projects.

Hoffman, Forrest [ORNL] (ORCID:0000000158024134)↗

Microgrid Integration with High Performance Computing Systems for Microreactor Operation

Multiple nuclear microreactor concepts are currently being developed across several sizes and fuel types with high performance computing (HPC) systems anticipated to be end-users of the power. Nuclear microreactors are small in size, portable, produce less than 10 MW electric, operate autonomously, and have a refueling interval of as many as 10 years. However, their load-follow is also generally limited to 10%/minute or worse whereas the power variance in HPC systems easily exceeds this constraint under normal operations. This study explores an approach that requires no load-follow from the microreactor but integrates the HPC system with a microgrid built from commercial-off-the-shelf components. Three typical HPC architectures are explored in the context of microgrid operation in this study. Components of power quality and transient response are empirically measured for five different HPC load-follow response levels using a self-contained mobile datacenter connected to the microgrid capable of integration with a nuclear microreactor.

microgrids↗

Performance evaluation of neutron noise analysis in detecting special nuclear materials

Reliable and rapid inspection techniques play a vital role in preventing illicit trafficking of special nuclear materials. Active interrogation systems using neutrons produced by portable, high-flux deuterium-deuterium or deuterium-tritium neutron generators are being actively developed as a secondary scanning tool for this purpose. In this study, a neutron noise analysis-based approach for detecting unshielded and shielded special nuclear materials by using a pulsed deuterium-tritium neutron generator was evaluated. Here, this approach analyzes the fluctuation of neutron counts. Its performance was quantified with regard to time-to-detection to achieve a minimum probability of detection of 99% and a probability of false alarm of less than 1% considering various amounts of special nuclear materials and different shielding configurations. It was demonstrated that this approach could detect 17 uranium slugs in 5 s given a neutron generator yield of 8.1 × 10 7 n/s. These slugs could be detected within a reasonable time frame (200 s) when they were shielded by 10.16 cm of high-density polyethylene. The results obtained using the neutron noise analysis approach were compared with those obtained using the commonly used differential die-away analysis technique, a sensitive technique for detecting the presence of fissile materials by utilizing the prompt fission neutrons produced when the source neutrons from a neutron generator are completely diminished. For example, the time to detect 2 unshielded uranium slugs was 2.1 s when using the differential die-away analysis technique; it increased to 93 s for the neutron noise analysis approach. Although the noise analysis-based approach exhibits an overall performance which is not as good as that of differential die-away, neutron noise provides an alternative method for effective detection of special nuclear materials.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Expanding the genetic toolset: using serine recombinases to integrate riboregulatory elements into industrially relevant microbial chassis

To realize the full potential of biomanufacturing, the breadth of industrial microbes used to consume diverse feedstock and generate bioproducts needs to expand. As such, portable tools are required that can be used by multiple hosts for straightforward genomic manipulation and precise gene expression. Here, we demonstrate the co-utilization of two synthetic biology tools to achieve these goals: cis-repressors (CRs) and serine recombinase-assisted genome engineering (SAGE). CRs are small, noncoding RNAs that are placed upstream of the target gene to modulate bacterial translation rates at varying, discrete levels. SAGE uses site-specific serine recombinases to catalyze highly efficient, unidirectional insertion of DNA into the chromosome of diverse organisms. We used SAGE to integrate a suite of CRs into the industrially relevant hosts Pseudomonas putida, Corynebacterium glutamicum, and Cupriavidus necator. Using a fluorescent reporter as a readout of CR functionality, we found that CR performance across these backgrounds was similar—providing a range of translational repression up to 100-fold. Overall, these results demonstrate the high portability of CRs across bacterial genetic backgrounds, which ideally can be used in future microbial engineering efforts pertinent to biomanufacturing.

59 BASIC BIOLOGICAL SCIENCES↗

Numerical Investigation of Fluid Flow and Space Charge in Liquid Argon Time Projection Chamber (LArTPC) Detectors

Overview This project focused on developing a high-fidelity numerical framework to simulate the multiphysics environment within Liquid Argon Time Projection Chamber (LArTPC) detectors. The primary objective was to characterize the complex interplay between ion transport, background fluid dynamics, and electric field distortions—a critical factor for the calibration and sensitivity of next-generation High Energy Physics experiments, such as DUNE. Technical Achievements The research successfully yielded a hybrid numerical space-charge solver utilizing a Cell-Centered Finite Volume Method (FVM) for ion transport coupled with a Finite Element Method (FEM) for electric potential. Key accomplishments include: • Verification & Validation: The 3-D solver was rigorously verified against 1-D analytical solutions, demonstrating high numerical accuracy in predicting space-charge-induced field deviations. • Field Distortion Analysis: 3D simulations revealed that space charge effects introduce significant non-uniformities in the electric field. Critically, the research identified that background LAr flow velocities, when comparable to ion drift velocities, markedly exacerbate these distortions. • Technology Transfer: The resulting source code and comprehensive user manuals were successfully transferred to collaborators at Fermilab, providing a portable computational tool for the broader scientific community. Challenges and Future Directions While the space-charge solver achieved all performance metrics, the integrated fluid dynamics modeling encountered convergence challenges stemming from the extreme 200-fold disparity in length scales between the detector's 37 mm inlet pipes and the 8-meter global domain. To address this, the project has identified a clear technical pivot toward Hierarchical Geometric Adaptive Mesh Refinement (HG-AMR). By implementing an h-type refinement strategy with hanging nodes, future iterations of this solver will be capable of resolving localized high-gradient inlet flows without the prohibitive computational costs of regular grids. This advancement, combined with data-driven uncertainty quantification based on MicroBooNE-style calibration, will enable the precise modeling of detector responses in large-scale cryogenic environments where direct measurement remains difficult. Impact The computational tools developed under this award provide a foundation for enhancing the energy resolution and spatial reconstruction of noble liquid detectors. By bridging the gap between theoretical fluid dynamics and experimental field calibration, this work supports the DOE’s mission to advance the frontiers of neutrino physics and dark matter detection.

42 ENGINEERING↗

Creating Apptainer Workflows with Docker-Compose-like Utilities

Creating Apptainer Workflows with Docker-Compose-like Utilities In this presentation, I will explore the utilization of a tool called process-compose, inspired by docker-compose, to create Apptainer-based services. This approach allows for easy deployment and management of fully containerized applications on High Performance Computing (HPC) systems without requiring elevated privileges. Benefits to the Ecosystem: By incorporating process-compose and Apptainer, I aim to address several key challenges in the HPC ecosystem: Simplified Workflow Management: Process-compose provides a user-friendly interface for defining and managing complex containerized application services, reducing the setup time and lowering the barrier to entry for new users. Enhanced Portability: Apptainer ensures that containerized applications can run consistently across different HPC environments, promoting greater portability and reducing compatibility issues. Process-compose is also a single binary that does not need to be installed by admin level users. Community Driven Solutions: This approach aligns with the goals of the High Performance Software Foundation (HPSF) to advance community-driven solutions. By sharing our experiences and insights, I hope to foster collaboration and innovation within the HPC community. Increased Productivity: The combination of process-compose and Apptainer streamlines the serve deployment process, allowing researchers and developers to focus more on their scientific work rather than the intricacies of system or service administration. Through this presentation, attendees will gain valuable insights into the practical implementation of containerized workflows on HPC systems, learn about the benefits of using process-compose and Apptainer, and understand how these tools can contribute to a more efficient HPC ecosystem.

97 - MATHEMATICS AND COMPUTING↗

Novel thermal energy storage component: Development, performance, and phase transition diagnosis

Thermal energy storage (TES) using phase change materials (PCMs) is a promising technology for capturing and storing excess thermal energy for later use. However, challenges such as poor heat transfer efficiency and a lack of modular, scalable designs have limited widespread adoption of TES in real-world applications. This study developed and evaluated modular brick-type and blade-type TES prototypes featuring an aluminum housing, an embedded serpentine coil for active or passive thermal exchange, and a cost-effective metal mesh to enhance PCM thermal conductivity. The blade-type TES achieved notable geometric efficiency, with a thickness-to-length ratio of 0.03 and a thickness-to-width ratio of 0.08, enabling highly compact and modular thermal storage suitable for space-constrained applications. The paper presents a detailed evaluation of the TES prototypes’ performance. The comparative analysis indicated that the TES prototypes provide a highly cost-effective, thermally optimized alternative for compact energy storage and load shifting. A novel diagnostic technique was also introduced: using a portable endoscope to capture real-time visualizations of PCM phase transitions inside the TES. This method provides critical insights into internal heat transfer mechanisms, identifies potential issues, and offers valuable support for optimizing the TES design and developing the control algorithm. Overall, the modular brick-type and blade-type TES designs demonstrated in this work provide a scalable, efficient, and economically viable solution for advancing TES across residential, commercial, and industrial sectors. The designs’ compact structure, enhanced thermal performance, and integrated diagnostic capabilities make them strong candidates for future deployment in energy-efficient systems.

Gao, Zhiming [ORNL] (ORCID:0000000271397995)↗

Evolution of the SLATE linear algebra library

SLATE (Software for Linear Algebra Targeting Exascale) is a distributed, dense linear algebra library targeting both CPU-only and GPU-accelerated systems, developed over the course of the Exascale Computing Project (ECP). While it began with several documents setting out its initial design, significant design changes occurred throughout its development. In some cases, these were anticipated: an early version used a simple consistency flag that was later replaced with a full-featured consistency protocol. In other cases, performance limitations and software and hardware changes prompted a redesign. Sequential communication tasks were parallelized; host-to-host MPI calls were replaced with GPU device-to-device MPI calls; more advanced algorithms such as Communication Avoiding LU and the Random Butterfly Transform (RBT) were introduced. Early choices that turned out to be cumbersome, error prone, or inflexible have been replaced with simpler, more intuitive, or more flexible designs. Applications have been a driving force, prompting a lighter weight queue class, nonuniform tile sizes, and more flexible MPI process grids. Of paramount importance has been building a portable library that works across several different GPU architectures – AMD, Intel, and NVIDIA – while keeping a clean and maintainable codebase. Here we explore the evolving design choices and their effects, both in terms of performance and software sustainability.

Gates, Mark↗

Error field detection and correction studies towards ITER operation

In magnetic fusion devices, error field (EF) sources, spurious magnetic field perturbations, need to be identified and corrected for safe and stable (disruption-free) tokamak operation. Within Work Package Tokamak Exploitation RT04, a series of studies have been carried out to test the portability of the novel non-disruptive method, designed and tested in DIII-D (Paz-Soldan et al 2022 Nucl. Fusion62 126007), and to perform an assessment of model-based EF control strategies towards their applicability in ITER. In this paper, the lessons learned, the physical mechanism behind the magnetic island healing, which relies on enhanced viscous torque that acts against the static electro-magnetic torque, and the main control achievements are reported, together with the first design of the asynchronous EF correction current/density controller for ITER.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Transforming Science Through Software: Improving While Delivering 100×

The U.S. Department of Energy (DOE) Exascale Computing Project (ECP) funded the development of new (and the transformation of important existing) applications, libraries, and tools that realized improvement in performance and capabilities of often 100 times or more on emerging exascale computers. This exceptional gain inspired the title of this special issue: Transforming Science through Software: Improving while delivering 100X. The term 100X refers to advancing capabilities in modeling, simulation, and analysis by a factor of 100 or more using some combination of new algorithms, optimization techniques, software libraries, and programming models, coupled with the next generation of hardware for high-performance computing (HPC). The papers in this issue share experiences with the practice and science of scientific software development, with an emphasis on developing a coherent, portable, and sustainable HPC software ecosystem for next-generation computational science. Finally, we hope to foster expanded community efforts related to the fundamental role of sustainable scientific software ecosystems in advancing the computing sciences.

97 MATHEMATICS AND COMPUTING↗

Elucidating Sodium-Sulfur Battery Chemistry using Operando Transmission X-ray Microscopy and X-ray Absorption Spectroscopy

Room-temperature sodium-sulfur (Na-S) batteries offer a beneficial and cost-effective solution for powering modern electric grids and devices. However, Na-S batteries suffer from performance decay during battery cycling due to the sluggish kinetics and polysulfide shuttling. Revealing the conversion mechanism of the Na-S chemistry is critical to designing high-capacity and stable Na-S batteries. In this study, we employed operando X-ray microscopy and operando X-ray absorption spectroscopy to elucidate the sulfur conversion in a carbon host using an ether-based electrolyte. We reveal the structural and chemical evolution of the sulfur cathode during cycling, which results in volume expansion and the dissolution of active material, ultimately affecting the battery performance. Our work provides fundamental insight into the sulfur reaction scheme, addressing the lingering challenges before establishing inexpensive Na-S batteries for broad applications, including portable electronics and large-scale energy storage.

Jagadeesan, Sathya Narayanan [SLAC National Accele↗

ChatBLAS: The First AI-Generated and Portable BLAS Library

We present ChatBLAS, the first AI-generated and portable Basic Linear Algebra Subprograms (BLAS) library on different CPU/GPU configurations. The purpose of this study is (i) to evaluate the capabilities of current large language models (LLMs) to generate a portable and HPC library for BLAS operations and (ii) to define the fundamental practices and criteria to interact with LLMs for HPC targets to elevate the trustworthiness and performance levels of the AI-generated HPC codes. The generated C/C++ codes must be highly optimized using device-specific solutions to reach high levels of performance. Additionally, these codes are very algorithm-dependent, thereby adding an extra dimension of complexity to this study. We used OpenAI’s LLM ChatGPT and focused on vector-vector BLAS level-1 operations. ChatBLAS can generate functional and correct codes, achieving high-trustworthiness levels, and can compete or even provide better performance against vendor libraries.

Valero Lara, Pedro↗

TRIM: AI Guided Random Number Generation for Resource-Constrained IoT Systems

Random numbers often serve as the backbone for many security solutions in diverse domains such as cryptography, side channel leakage prevention, and moving target defense. However, generating true random numbers requires a physical source of entropy (e.g. hardware, quantum, environmental phenomenon) making it difficult to realize at a large scale and at a low cost. On the flip side, pseudorandom number generators (easy to implement) following a specific distribution (e.g. Gaussian) can be easily compromised given a sufficient amount of traces. In this work, we have developed a machine learning-guided generative approach that can be used to create portable, resource-efficient, and cost-effective random number generators with high throughput and true randomness characteristics. We implement the proposed approach as a highly parameterized framework and perform extensive evaluation for different settings. The framework was able to learn from true random sources such as irrational numbers and environmental audio noise and imitate those sources towards generating new good quality random numbers on demand. We have generated more than 1 billion bits and observed robust performance in terms of true randomness metrics obtained from NIST SP 800-22 and FIPS 140-1 randomness test suites achieving a throughput of up to 142.85 Mbps. Compared to the state-of-the-art (SOTA) technique, the iso-cost setup of our framework can achieve more than 500 Mbps in a distributed setting. We have evaluated the efficacy of running the true randomness imitation AI models on target edge devices such as Raspberry Pi 4 (Model B), Nvidia Jetson Nano, Nvidia Jetson Orin Nano and Nvidia Jetson Xavier. We have also looked at the security of the TRIM framework itself against different adversarial threat models.

Cybersecurity↗

PETSc/TAO Users Manual Revision 3.22

This manual describes the use of the Portable, Extensible Toolkit for Scientific Computation (PETSc) and the Toolkit for Advanced Optimization (TAO) for the numerical solution of partial differential equations (PDEs) and related problems on high-performance computers. PETSc/TAO is a suite of data structures and routines that provide the building blocks for implementing large-scale application codes on parallel (and serial) computers. PETSc uses the MPI standard for all distributed memory communication.

97 MATHEMATICS AND COMPUTING↗

PETSc/TAO Users Manual Revision 3.23

This manual describes the use of the Portable, Extensible Toolkit for Scientific Computation (PETSc) and the Toolkit for Advanced Optimization (TAO) for the numerical solution of partial differential equations (PDEs) and related problems on high-performance computers. PETSc/TAO is a suite of data structures and routines that provide the building blocks for implementing large-scale application codes on parallel (and serial) computers. PETSc uses the MPI standard for all distributed memory communication.

97 MATHEMATICS AND COMPUTING↗

PETSc/TAO Users Manual Revision 3.24

This manual describes the use of the Portable, Extensible Toolkit for Scientific Computation (PETSc) and the Toolkit for Advanced Optimization (TAO) for the numerical solution of partial differential equations (PDEs) and related problems on high-performance computers. PETSc/TAO is a suite of data structures and routines that provide the building blocks for implementing large-scale application codes on parallel (and serial) computers. PETSc uses the MPI standard for all distributed memory communication.

97 MATHEMATICS AND COMPUTING↗