Search NASA⌕ Search

SEARCH · Search NASA

Results for “performance optimization”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 487 records · Page 27

Real-Time Lifetime Prediction of Semiconductor Devices Using Hardware-in-the-Loop

This paper presents a unique approach to enable real-time lifespan prediction of semiconductor power modules using a Hardware-in-the-Loop (HIL) system. By integrating the module's overall loss characteristics-specifically switching and conduction losses-with a thermoelectric model of the thermal management system, this research demonstrates that the model can dynamically estimates the junction temperature profile of the semiconductor devices in response to a changing torque demand profile for the motor drive system. This capability enables continuous monitoring of the module's operational time and cumulative stress induced on the devices to compute accumulated remaining lifetime or time-to-failure (TTF). This study provides an architectural framework for the HIL system with high-fidelity component models of multiple physical domains, allowing simulation of dynamic behaviors of a closely-coupled motor drive system. The advanced real-time computation and measurement functionalities of the HIL system allow for both dynamic lifetime calculations based on simulated data and aggregate lifetime predictions utilizing historical data. Moreover, this paper details an algorithm that not only computes cumulative damage but also synthesizes these data into a comprehensive aggregated lifetime metric. This methodology can enhance the maintenance scheduling strategies and operational reliability of semiconductor devices in critical applications, ultimately extending their service life while optimizing performance.

hardware-in-the-loop (HIL)↗

Distributed Multi-GPU Community Detection on Exascale Computing Platforms

Community detection is a fundamental operation in graph mining, and by uncovering hidden structures and patterns within complex systems it helps solve fundamental problems pertaining to social networks, such as information diffusion, epidemics, and recommender systems. Scaling graph algorithms for massive networks becomes challenging on modern distributed-memory multi-GPU (Graphics Processing Unit) systems due to limitations such as irregular memory access patterns, load imbalances, higher communication-computation ratios, and cross-platform support. We present a novel algorithm HiPDPL-GPU (distributed parallel Louvain) to address these challenges. We conduct experiments involving different partitioning techniques to achieve optimized performance of HiPDPL-GPU on the two largest supercomputers: Frontier and Summit. Remarkably, HiPDPL-GPU processes a graph with 4.2 billion edges in less than 3 minutes using 1024 GPUs. Qualitatively performance of HiPDPL-GPU is similar or better compared to other state-of-the-art CPU- and GPU-based implementations. While prior GPU implementations have predominantly employed CUDA, our first-of-its-kind implementation for community detection is cross-platform, accommodating both AMD and NVIDIA GPUs.

graph algorithms, high performance comptuing↗

Automatic Extraction of Network Configurations for Realistic Simulation and Validation

Popular HPC network interconnection simulators such as SST Macro provide a variety of configurable parameters to explore the design space of hardware components such as network links and switches. While such knobs provide flexibility to explore design trade-offs for novel hardware, manually configuring simulations for existing hardware to focus on topology exploration can be cumbersome and error-prone, leading to widely inaccurate simulations. This challenge is compounded when specifications of various (proprietary) technologies are not readily available or are intentionally omitted. In this work, we provide a methodology to automatically tune the simulation configuration of the multiple network models running within SST Macro using Bayesian optimization. We perform this optimization in the context of multiple messaging regimes (i.e., small to large and latency to bandwidth-bound messages) and provide a detailed analysis of the simulation error for four systems. With our automated framework, we achieve a 5x improvement in accuracy over best-effort configurations based on available hardware specifications.

Suetterlein, Joshua D.↗

MDLoader: A Hybrid Model-Driven Data Loader for Distributed Graph Neural Network Training

Scalable data management is essential for processing large scientific dataset on HPC platforms for distributed deep learning. In-memory distributed storage is preferred for its speed, enabling rapid, random, and frequent data access required by stochastic optimizers. Processes use one-sided or collective communication to fetch remote data, with optimal performance depending on (i) dataset characteristics, (ii) training scale, and (iii) interconnection network. Empirical analysis shows collective communication excels with larger mini-batch sizes and/or fewer processes, whereas one-sided communication outperforms at larger scales. We propose MDLoader, a hybrid in-memory data loader for distributed graph neural network training. MDLoader features a model-driven performance estimator that dynamically selects between one-sided and collective communication at the beginning of training using Tree of Parzen Estimators (TPE). Evaluations on NERSC Perlmutter and OLCF Summit show MDLoader outperforms single-backend loaders by up to 2.83 × and predicts the suitable communication method with 96.3% (Perlmutter) and 94.3% (Summit) success rate.

Bae, Jonghyun↗

An Experimental Setup for Mechanical Vibration Analysis Using VLC

This study explores the potential applications across various domains, including earthquake detection and warning systems, where the system’s sensitivity to ground vibrations can contribute to early seismic event detection. Additionally, the study paves the way of developing applications of VLC/T in mechanical vibration and stability analysis of engines and platforms, offering insights into structural integrity and performance optimization. These multifaceted applications underscore the adaptability and potential of VLC/T systems in diverse fields, heralding advancements in sensing, communication, and security technologies. To achieve this, in this study, Peak to Average Power Ratio (PAPR) is proposed to represent the impact of mechanical shocks and vibrations generated by several weights dropped onto the platform with which the receiver is fixed. Even though non-contact measurement methodology is preferred for various reasons, the proposed measurement campaign obtains the data in contact form; however, the system and signal model proposed in this study could easily be extended into non-contact form. Considering the fact that the proposed measurement campaign employs off-the-shelf products, it is cost-effective and very scalable.

Yilmaz, Ahmet Mucahit↗

Memory-Aware External Facelist Calculation: A Data-Parallel Atomic Hash Counting Approach

Unstructured volumetric meshes serve as fundamental data representations in various scientific simulations and analyses. They play a crucial role in representing complex computational domains and are essential for important numerical techniques, such as finite element analysis. Whenever such a mesh is read from a file, streamed in-situ, or generated by algorithms, scientific visualization libraries rely on calculating the external surface of a geometry, named “external facelist”, to produce a polygonal mesh for rendering. Consequently, external facelist calculation has become one of the most widely used algorithms in the scientific visualization domain, necessitating optimal performance. In this paper, we explore relevant work on external facelist calculation algorithms in two common visualization libraries, VTK and Viskores, assess their performance and memory constraints, and introduce a novel memory-aware external facelist calculation algorithm employing an atomic hash counting approach. This algorithm fully leverages Viskores' data-parallel primitive operations, facilitating its execution across diverse many-core architectures. Our algorithm features the lowest memory footprint on the GPU and the second-lowest on the CPU among all evaluated methods, and it also delivers the fastest performance on both CPU and GPU. It has been made available under an open-source license in the VTK and Viskores visualization systems.

Tsalikis, Spiros [Kitware] (ORCID:0000000151137195↗

Strangers in a foreign land: ‘Yeastizing’ plant enzymes

Abstract Expressing plant metabolic pathways in microbial platforms is an efficient, cost‐effective solution for producing many desired plant compounds. As eukaryotic organisms, yeasts are often the preferred platform. However, expression of plant enzymes in a yeast frequently leads to failure because the enzymes are poorly adapted to the foreign yeast cellular environment. Here, we first summarize the current engineering approaches for optimizing performance of plant enzymes in yeast. A critical limitation of these approaches is that they are labour‐intensive and must be customized for each individual enzyme, which significantly hinders the establishment of plant pathways in cellular factories. In response to this challenge, we propose the development of a cost‐effective computational pipeline to redesign plant enzymes for better adaptation to the yeast cellular milieu. This proposition is underpinned by compelling evidence that plant and yeast enzymes exhibit distinct sequence features that are generalizable across enzyme families. Consequently, we introduce a data‐driven machine learning framework designed to extract ‘yeastizing’ rules from natural protein sequence variations, which can be broadly applied to all enzymes. Additionally, we discuss the potential to integrate the machine learning model into a full design‐build‐test cycle.

59 BASIC BIOLOGICAL SCIENCES↗

Sixteen multiple-amplifier sensing charge-coupled devices and characterization techniques targeting the next generation of astronomical instruments

We present a candidate sensor for future spectroscopic applications, such as a Stage-5 Spectroscopic Survey Experiment or the Habitable Worlds Observatory. This type of charge-coupled device (CCD) sensor features multiple in-line amplifiers at its output stage allowing multiple measurements of the same charge packet, either in each amplifier or in the different amplifiers. Recently, the operation of an eight-amplifier sensor has been experimentally demonstrated, and we present the operation of a 16-amplifier sensor. This new sensor enables a noise level of ∼1 erms− with a single sample per amplifier. In addition, it is shown that sub-electron noise can be achieved using multiple samples per amplifier. In addition to demonstrating the performance of the 16-amplifier sensor, we aim to create a framework for future analysis and performance optimization of this type of detectors. New models and techniques are presented to characterize specific parameters, which are absent in conventional CCDs and Skipper CCDs: charge transfer between amplifiers and independent and common noise in the amplifiers and their processing.

16 multiple-amplifer sensing CCD (MAS-CCD)↗

OpenARC

OpenARC is an open-sourced, very High-Level Intermediate Representation (HLIR)-based, extensible compiler framework, where various performance optimizations, traceability mechanisms, fault tolerance techniques, etc., can be built for better debuggability/performance/resilience on the complex accelerator computing. OpenARC is the first OpenACC compiler supporting Altera FPGAs, in addition to NVIDIA GPUs, AMD GPUs, and Intel Xeon Phis.

Lee, Seyong [Oak Ridge National Laboratory (ORNL),↗

Vedizar Fingerprinter

SAND2025-03289O Vedizar Fingerprinter simplifies the process of identifying devices on a network by analyzing traffic data. It uses a unique library to recognize different devices, making it easier for users to understand what is happening on their networks. This software is ideal for IT and operational technology environments, helping organizations monitor their networks effectively. By saving results in a database, it allows for easy access and review of device information. Users can enhance their network security and optimize performance without needing specialized hardware or technical expertise. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Jacobellis, John [Sandia National Lab. (SNL-CA), L↗

SpectraCodec: A Hilbert curve-based method for encoding metadata in mass spectra for machine learning applications (SpectraCodec) v1

Machine learning approaches to mass spectrometry (MS) data analysis require structured metadata for optimal performance. However, current MS file formats necessitate external metadata sources, creating integration challenges that impede analytical workflows. Here, we present a novel approach for encoding metadata directly within mzML files using one-hot encoding of ASCII characters mapped via Hilbert space-filling curves. This strategy embeds metadata in the first spectrum's m/z-intensity space, ensuring persistence with the primary data, eliminating the need for external metadata files, and maintaining compatibility with existing MS software. We demonstrate that the Hilbert curve mapping efficiently utilizes the two-dimensional spectral space while maintaining robust data recovery. This method offers a practical solution for machine learning applications in mass spectrometry by ensuring metadata and spectral data remain unified through all stages of analysis.

Bowen, Benjamin [Lawrence Berkeley National Labora↗

Pumpchart v1.0.0

Pumpchart is a graphical tool that plots the state of a hydraulic system overlayed on a pump curve and system curve. It assists with determining the optimal performance of the system. One advantage of Pumpchart over other softwares is that it is integrated with the Grafana dashboarding software, so that data can be plotted in real-time for instant operator feedback. It is intended to be used at NERSC to monitor the performance of our cooling water pumps.

Venture, Nicholas [Lawrence Berkeley National Labo↗

STATUS OF THE SECOND INTERACTION REGION DESIGN FOR ELECTRON-ION COLLIDER

Provisions are being made in the Electron Ion Collider (EIC) design for future installation of a second Interaction Region (IR), in addition to the day-one primary IR. The envisioned location for the second IR is the existing experi- mental hall at RHIC IP8. It is designed to work with the same beam energy combinations as the first IR, covering a full range of the center-of-mass energy of ?20 GeV to ?140 GeV. The goal of the second IR is to complement the first IR, and to improve the detection of scattered particles with magnetic rigidities similar to those of the ion beam. To achieve this, the second IR hadron beamline features a secondary focus in the forward ion direction. The design of the second IR is still evolving. This paper reports the current status of its pa- rameters, magnet layout, and beam dynamics and discusses the ongoing improvements being made to ensure its optimal performance.

Gamage, B.↗

BOPTest As a Platform for Building Controls and Grid-Interactive Buildings Workforce Training

Building automation and controls are becoming increasingly complex with the emergence of Grid Integrated Efficient Buildings (GEBs) as well as new highly efficient sequences of operation and data-driven control schemes. However, there remains a significant gap in hands-on training opportunities for building operators and technicians to gain practical experience with advanced control systems in a low-risk environment. This paper presents BOPTEST (Building Optimization Performance Test) as a suitable platform for workforce training in building controls and GEB technologies. BOPTEST provides a suite of standardized building simulation test cases with a REST API, real-time control interfaces through BACnet, semantic models connecting users to building data, and built-in calculation of control metrics and performance indicators. The platform enables trainees to interact with virtual buildings using industry-standard protocols while learning how to implement and innovate control strategies. The training platform is designed to offer a structured and interactive learning experience for building engineers, helping them effectively develop, learn, and retain skills in fault identification, troubleshooting, and correction. The workflow is divided into three main phases: 1) Setup, 2) Exercise, and 3) Review, each comprising specific activities performed by either the instructor or the student. Initial pilot training sessions have yielded positive feedback from instructors and participants and demonstrates that BOPTEST effectively fills an industry need for a low-risk training resource via simulation of real building control systems, allowing trainees to gain practical experience before working in the field. The platform's ability to provide immediate performance feedback while maintaining familiar industry interfaces makes it particularly suitable for workforce development programs. This work provides a replicable model for leveraging building simulation in control education and training.

Paul, Lazlo↗

STATUS OF THE SECOND INTERACTION REGION DESIGN FOR ELECTRON-ION COLLIDER

Provisions are being made in the Electron Ion Collider (EIC) design for future installation of a second Interaction Region (IR), in addition to the day-one primary IR. The envisioned location for the second IR is the existing experi- mental hall at RHIC IP8. It is designed to work with the same beam energy combinations as the first IR, covering a full range of the center-of-mass energy of ?20 GeV to ?140 GeV. The goal of the second IR is to complement the first IR, and to improve the detection of scattered particles with magnetic rigidities similar to those of the ion beam. To achieve this, the second IR hadron beamline features a secondary focus in the forward ion direction. The design of the second IR is still evolving. This paper reports the current status of its pa- rameters, magnet layout, and beam dynamics and discusses the ongoing improvements being made to ensure its optimal performance.

Gamage, B.↗

Evaluation of the Stability of Edge-Passivated CZTS Radiation Detectors

CdZnTeSe (CZTS) is a next gen replacement room temperature radiation detector that solves common issues with modern CZT detectors. • Surface states can trap radiation induced carriers and can act as leakage current pathways requiring detectors to passivated for optimal performance. • In this work, we seek to identify the best passivation technique for CZTS detectors.

KLEPPINGER, JOSHUA↗

An Analytical Tool to Evaluate Defect Thermodynamics of (La,Ba)Fe1-xMxO3-δ Perovskites for Solid-Oxide Cell Applications

A modeling tool of the defect thermodynamics of (La,Ba)Fe1-xMxO3-δ perovskites which includes energetic information about oxygen vacancy formation, hydration, hydride formation, and charge disproportionation reactions has been developed1. This tool incorporates defect energies and entropies expressed as sixth order polynomial functions to allow refinements of the defect reaction equilibrium constants in the thermodynamic analysis. Calculation of (La,Ba)Fe1-xMxO3-δ Brouwer diagrams as a function of pO2/pH2O and pH2/H2O in a range of temperatures of interest is facilitated by this modeling tool. The results obtained can provide direct guidance how the electronic and ionic defect concentrations of the triple conducting perovskite materials can be used to optimize performance of solid oxide cells for energy applications. The impact of magnetic and electronic structures of the perovskites on the defect reaction energies and entropies as obtained from density function theory modeling and the role played by hydride defect species will also be discussed.

Lee, Yueh-Lin↗

Evaluate Synergies of Using Hydrothermal Liquefaction and Anerobic Digestion Treatment Technologies for Wastewater Resource Recovery Facilities (CRADA 516 Final Report)

The research focuses on utilizing a new anaerobic digestion (AD) configuration to treat the aqueous by-product generated by hydrothermal liquefaction (HTL) of sewage sludge. This report found that for Anaerobic Digestion for HTL By-product, Anaerobic biofilms can degrade some HTL wastewater contaminants, but co-digestion is essential to address nutrient deficiencies and optimize performance. Without AD, toxicity of HTL aqueous streams may limit broader adoption in wastewater treatment plants (WWTPs). Great Lakes Water Authority (GLWA) used an innovative reactor design, involving a dynamic membrane anaerobic bioreactor to promote biofilm growth, improving contaminant degradation. The tree-like structure inside the reactor supports biofilm development with recirculation enhancing microbial activity. Overall, a 70% chemical oxygen demand (COD) removal was achieved, although nutrient supplementation is required for stability. The reactor achieved a diverse microbial community, including methanogens and bacteria capable of degrading phenols and aromatics.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗