Search NASA⌕ Search

SEARCH · Search NASA

Results for “Performance Tuning”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 163 records · Page 9

Predicting the von Neumann entanglement entropy using a graph neural network

Calculating the von Neumann entanglement entropy from experimental data is challenging due to its dependence on the complete wavefunction, forcing reliance on approximations such as classical mutual information (MI). We propose a machine learning approach using a graph neural network to predict the von Neumann entropy directly from experimentally accessible bitstrings. We test this approach on a Rydberg ladder system and achieve a mean absolute error of $3.6\,\times 10^{-3}$ when evaluating within the training range on a dataset with entropy values ranging from 0 to 1.9. The model achieves a mean absolute percentage error of 1.44% and outperforms MI-based bounds. When tested beyond the training range, the model maintains reasonable accuracy. Furthermore, we demonstrate that fine-tuning the model with small datasets significantly improves performance on data outside the original training range.

graph neural networks↗

Segmentation Model Distillation [Poster]

The process of training object detection (OD) or image segmentation model requires both a substantial amount of data and technical knowledge, which often creates challenges in applying these types of models to their full potential. In order to streamline the process of developing these models, we propose a new pipeline where a foundation model assists in the dataset generation. Then this resulting dataset is used to fine-tune a fast light-weight model to perform the custom segmentation or OD. This resulting model is also fit for real-time image segmentation, such as in a video stream.

97 MATHEMATICS AND COMPUTING↗

Machine Learning-Based Process Control for Injection Molding of Recycled Polypropylene

The increased interest in artificial intelligence in manufacturing has driven the adoption of machine learning to optimize processes and improve efficiency. A key challenge in injection molding is the variability of recycled materials, which affects part quality and processing stability. This study presents a novel closed-loop process control approach for injection molding, leveraging machine learning to adaptively predict processing inputs and quality outcomes. The methodology was tested on five blends of recycled polypropylene (rPP), using artificial neural networks (ANNs), linear regression, and polynomial regression to model the relationships between material properties and process parameters. The dataset was split 80/20 into training and testing sets. The ANN model was implemented using TensorFlow and Keras, with six hidden layers of 32 neurons per layer, ReLU activation, and an Adam optimizer. Empirical tuning and early stopping were used to optimize performance and prevent overfitting. Predictions were evaluated based on mean absolute error (MAE), mean squared error (MSE), and percentage error. The results showed that yield stress, ultimate elongation, and part weight were accurately predicted within a 5% error for linear and polynomial regression models and within a 10% error for the ANN. However, modulus predictions were less reliable, with errors of ~11% for ANN and linear regression and ~40% for polynomial regression, reflecting the inherent variability of this property in rPP blends. Predictions of processing inputs had errors ranging from 3% to 25%, depending on the model and response variable. No single modeling approach was consistently superior across all responses, highlighting the complexity of the relationship between material properties, process parameters, and quality metrics. Overall, the work demonstrates that closed-loop process control, powered by machine learning, can effectively predict key quality parameters in injection molding of recycled materials. The proposed approach can improve process stability and material utilization, facilitating increased adoption of sustainable materials.

Krantz, Joshua↗

Text Mining for Process–Structure–Properties Relationships in Metals

With the advent of large language models (LLMs), the vast unstructured text within millions of academic papers is increasingly accessible for materials discovery—although significant challenges remain. While LLMs offer promising few- and zero-shot learning capabilities, particularly valuable in the materials domain where expert annotations are scarce, general-purpose LLMs often fail to address key materials-specific queries without further adaptation. To bridge this gap, fine-tuning LLMs on human-labeled data is essential for effective structured knowledge extraction (Liu in The Importance of Human-Labeled Data in the Era of LLMs, 2023). Here, in this study, we introduce a novel annotation schema designed to extract generic process–structure–properties relationships from scientific literature. We demonstrate the utility of this approach using a dataset of 128 abstracts, with annotations drawn from two distinct domains: high-temperature materials (Domain I) and uncertainty quantification in simulating materials microstructure (Domain II). Initially, we developed a conditional random field (CRF) model based on MatBERT—a domain-specific BERT variant—and evaluated its performance on Domain I. Subsequently, we compared this model with a fine-tuned LLM (GPT-4o from OpenAI) under identical conditions. Our results indicate that fine-tuning LLMs can significantly improve entity extraction performance over the BERT-CRF baseline on Domain I. However, when additional examples from Domain II were incorporated, the performance of the BERT-CRF model became comparable to that of the GPT-4o model. These findings underscore the potential of our schema for structured knowledge extraction and highlight the complementary strengths of both modeling approaches.

Materials science↗

Fermilab Booster loss modelling and rebalancing using Bayesian methods

Fermilab Booster is being upgraded for the PIP-II project to support 20Hz ramp rate at higher intensities. Loss trip limits determine the achievable peak power. To meet PIP-II requirements, losses need to be halved as compared to current levels. Losses primarily occur at injection and transition crossing, with both gradually increasing and threshold-like intensity-dependent behaviors. The existing simulation models are not yet good enough for quantitative loss predictions. In practice, it will be necessary to tune up the Booster using iterative methods and operator intuition. In this paper we present an effort to systematically model Booster losses using active learning (Bayesian exploration) techniques, and subsequently to rebalance them for higher trip limit margins. We first created several sets of spatially and temporally isolated orbit and optics knobs, and trained Gaussian process models for each beam loss monitor as well as beam current. This is a complex task due to safety and timing requirements – we discuss mitigations such as uncertainty constraints and approximate fitting. Once models are stable, we perform large-scale single and multi-objective tuning using scalarized objectives made up of critical beam loss locations. Our results demonstrate significant rebalancing of losses, increasing trip margins, as well as an overall improvement in beam transmission efficiency. We are exploring how to combine existing simulations with experimental data and automate the collection procedure so that more advanced surrogate models can be created over time.

Kuklev, Nikita [Fermilab]↗

Analyzing and Exploring Training Recipes for Large-Scale Transformer-Based Weather Prediction

Abstract The rapid rise of deep learning (DL) in numerical weather prediction (NWP) has led to a proliferation of models which forecast atmospheric variables with comparable or superior skill than traditional physics-based NWP. However, among these leading DL models, there is a wide variance in both the training settings and architecture used. Further, the lack of thorough ablation studies makes it hard to discern which components are most critical to success. In this work, we show that it is possible to attain high forecast skill even with relatively off-the-shelf architectures, simple training procedures, and moderate compute budgets. Specifically, we train a minimally modified Swin Transformer V2 (SwinV2) on ERA5 data and find that it attains superior skill in terms of mean-square errors of deterministic forecasts when compared against the European Centre for Medium-Range Weather Forecasts’ Integrated Forecasting System (IFS). Almost all DL–NWP systems share a core set of hyperparameters and design decisions. To aid and expedite future DL–NWP research, we present an in-depth, systematic exploration of different loss functions, model sizes and depths, patch sizes, and multistep training objectives. We also examine the model performance with metrics beyond the typical accuracy (ACC) and RMSE and investigate how the performance scales with model size. Through our open-source code, scoring pipelines, and models, we share our findings on key aspects of the training pipeline. These ablations reduce the necessity for expensive hyperparameter tuning and lower the barrier to entry for future DL–NWP research. Significance Statement This study investigates the potential of using large-scale transformer-based models for weather prediction, showing that it is possible to achieve high forecast accuracy with simpler, off-the-shelf architectures. By training a minimally modified SwinV2 transformer on ERA5 data, we show that the model achieves competitive forecast skill in terms of mean-square error for key variables, outperforming the European Centre for Medium-Range Weather Forecasts’ Integrated Forecasting System (IFS) at all lead times. Our findings suggest that effective training strategies, such as multistep fine-tuning and channel-weighted losses, significantly enhance the model’s performance. However, we also highlight that these improvements come with trade-offs in other areas, such as ensemble spread and high-frequency spatial detail. This work highlights the promise of deep learning in improving weather forecasts, which could lead to better preparedness and response to weather events, ultimately benefiting society by providing more reliable weather predictions.

Willard, Jared D. [Lawrence Berkeley National Labo↗

Intrinsic scintillation performance & europium concentration effects in RbSr 2 I 5 and RbSr 2 Br 5 scintillators

Scintillators play crucial roles in homeland security applications like gamma ray spectroscopy and high energy X-ray radiography. For promising new scintillators, fine-tuning the luminescent dopant concentration is one avenue to further improve their performance and tailor their properties. In this work, the effects of europium dopant concentrations on the crystal growth, luminescence and scintillation properties of RbSr 2 Br 5 and RbSr 2 I 5 crystals was investigated. Nine transparent 7 mm diameter single crystals were grown via the Vertical Bridgman method. Here, the optical band gap of RbSr 2 Br 5 was 5.9 eV and that of RbSr 2 I 5 was 4.7 eV. High scintillation performance was achieved with a relatively low europium concentration of 1 mol%. For both RbSr 2 Br 5 :Eu and RbSr 2 I 5 :Eu crystals, light yield of 60–90,000 ph/MeV, energy resolution 2.8–4.0 % at 662 keV, and X-ray afterglow 0.79–1.5 % at 2 ms were obtained.

36 MATERIALS SCIENCE↗

Upgrade to Fixed and Translating Scintillation-Based Loss Detector System in the Fermilab Drift Tube Linac

The closed-off structure of the Fermilab Drift Tube Linac precludes a robust array of instrumentation from directly monitoring the H- beam that is accelerated from 750 keV to 116 MeV. To improve beam tuning and op-erational assessment of Drift Tube Linac performance, scintillator-based loss monitors were previously installed along the exterior of the first two accelerating cavities to assess low energy beam losses. Here we present a recent upgrade to the loss monitor system, including significant improvements in analog signal processing to address baseline-interfering noise; digitization of the signals to enable regular operational use and tuning; and a new remote operation upgrade of the translating loss monitor with precise positioning of the loss monitor along its nine-foot track. Data from the fixed and translating de-tectors collected under varying beam conditions validate the utility of the upgrade.

Chen, Erin V. [Fermilab] (ORCID:0000000215467899)↗

Performance-Aligned LLMs for Generating Fast HPC Code

Optimizing scientific software is a difficult task because codebases are often large and complex, and performance can depend upon several factors including the algorithm, its implementation, and hardware among others. Causes of poor performance can originate from disparate sources and be difficult to diagnose. Recent years have seen a multitude of work that use large language models (LLMs) to assist in software development tasks. However, these tools are trained to model the distribution of code as text, and are not specifically designed to understand performance aspects of code. In this work, we introduce a reinforcement learning based methodology to align the outputs of code LLMs with performance. This allows us to build upon the current code modeling capabilities of LLMs and extend them to generate better performing code. Here, we demonstrate that our fine-tuned model improves the expected speedup of generated code over base models for a set of benchmark tasks from 0.9 to 1.6 for serial code and 1.9 to 4.5 for OpenMP parallel code.

Computer science↗

Transformers and Long Short-Term Memory Transfer Learning for GenIV Reactor Temperature Time Series Forecasting

Automated monitoring of the coolant temperature can enable autonomous operation of generation IV reactors (GenIV), thus reducing their operating and maintenance costs. Automation can be accomplished with machine learning (ML) models trained on historical sensor data. However, the performance of ML usually depends on the availability of large amount of training data, which is difficult to obtain for GenIV, as this technology is still under development. We propose the use of transfer learning (TL), which involves utilizing knowledge across different domains, to compensate for this lack of training data. TL can be used to create pre-trained ML models with data from small-scale research facilities, which can then be fine-tuned to monitor GenIV reactors. In this work, we develop pre-trained Transformer and long short-term memory (LSTM) networks by training them on temperature measurements from thermal hydraulic flow loops operating with water and Galinstan fluids at room temperature at Argonne National Laboratory. The pre-trained models are then fine-tuned and re-trained with minimal additional data to perform predictions of the time series of high temperature measurements obtained from the Engineering Test Unit (ETU) at Kairos Power. The performance of the LSTM and Transformer networks is investigated by varying the size of the lookback window and forecast horizon. The results of this study show that LSTM networks have lower prediction errors than Transformers, but LSTM errors increase more rapidly with increasing lookback window size and forecast horizon compared to the Transformer errors.

LSTM↗

Surrogate-Based Autotuning for Randomized Sketching Algorithms in Regression Problems

Algorithms from Randomized Numerical Linear Algebra (RandNLA) are known to be effective in handling high-dimensional computational problems, providing high-quality empirical performance as well as strong probabilistic guarantees. However, their practical application is complicated by the fact that the user needs to set various algorithm-specific tuning parameters which are different from those used in traditional NLA. This paper demonstrates how a surrogate-based autotuning approach can be used to address fundamental problems of parameter selection in RandNLA algorithms. In particular, we provide a detailed investigation of surrogate-based autotuning for sketch-and-precondition (SAP)-based randomized least squares methods, which have been one of the great success stories in modern RandNLA. Empirical results show that our surrogate-based autotuning approach can achieve near-optimal performance with much less tuning cost than a random search (up to about 7.6x fewer trials of different parameter configurations). Moreover, while our experiments focus on least squares, our results demonstrate a general-purpose autotuning pipeline applicable to any kind of RandNLA algorithm.

Cho, Younghyun↗

Comparative Analysis of Model Predictive Control and MPC-Informed Rule-Based Control for Thermal Storage Operation in Ultra-Low Temperature 4th Generation District Heating Networks

The integration of thermal storage and heat pumps in district heating networks (DHNs) can significantly enhance operational flexibility and energy efficiency; however, the practical deployment of advanced control strategies is often hindered by forecasting requirements and computational complexity. This study presents a comparative analysis of thermal storage control strategies in an ultra-low-temperature fourth-generation DHN, focusing on the development of a simplified rule-based control (RBC) explicitly informed by Model Predictive Control (MPC) behavior. The proposed methodology systematically analyzes the charging and discharging decisions of an MPC-controlled system under ideal forecasting conditions and extracts recurrent control patterns as a function of key system variables, including outdoor temperature, thermal demand, and electricity price. These patterns are translated into a set of structured time- and condition-based rules, resulting in an MPC-informed RBC that embeds predictive insights while preserving implementation simplicity and operational transparency. The approach is validated on a realistic mixed-use urban district in Denver, Colorado, USA, equipped with a centralized air-source heat pump, distributed water-to-water heat pumps, and a central thermal storage unit. Results show that the tuned RBC attains approximately 96% of ideal MPC economic performance (-27% of costs), preserves values of technical and environmental indicators (reduction only of 2-3%), and substantially reduces complexity. Sensitivity analyses further demonstrate the robustness of the RBC under varying operational conditions (i.e., ambient temperature, electricity price). Overall, the study demonstrates that MPC-informed rule-based control represents an effective trade-off between control performance and real-world applicability, enabling the integration of additional system components while maintaining simplicity, robustness, and ease of implementation.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Parallel computing for power system climate resiliency: Solving a large-scale stochastic capacity expansion problem with mpi-sppy

Here we propose a nodal stochastic generation and transmission expansion planning model that incorporates the output from high-resolution global climate models through load and generation availability scenarios. We implement our model in Pyomo and perform computational studies on a realistically-sized test case of the California electric grid in a high performance computing environment. We propose model reformulations and algorithm tuning to efficiently solve this large problem using a variant of the Progressive Hedging Algorithm. We utilize the parallelization capabilities and overall versatility of mpi-sppy, exploiting its hub-and-spoke architecture to concurrently obtain inner and outer bounds on an optimal expansion plan. Initial results show that instances with 360 representative days on a system with over 8,000 buses can be solved to within 5% of optimality in under 4 h of wall clock time, a first step towards solving a large-scale power system expansion planning problem across a wide range of climate-informed operational scenarios.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Unraveling the Structural Sensitivity of Metal Catalysts in Ethylene Hydroformylation: Insights from Theory and Experiments

Here, in this study, we combined experimental and theoretical methods to investigate the structural sensitivity of metal catalysts in the ethylene hydroformylation reaction. Among Rh, Pt, Ir, Ni, Au, Ag, Pd, and Cu catalysts studied using experimental and theoretical methods, Rh showed the highest selectivity toward the C-C coupling product from CO and C 2 H 5 (i.e., C 2 H 5 CHO). The results from DFT and microkinetic simulations revealed that the activation energy barrier for C-C coupling is lowest on the Rh nanocluster, which explains the experimentally observed highest C 2 H 5 CHO selectivity on the Rh catalyst. Furthermore, DFT results demonstrated that the sites located on the flat surfaces of nanoparticles primarily promote the hydrogenation reaction, leading to the formation of undesired C 2 H 6 . In contrast, undercoordinated edge and corner sites of the nanocluster promote the C-C coupling reaction. Thus, our results illustrate that the selectivity toward C 3 oxygenates in ethylene hydroformylation reaction can be steered by tuning the size of Rh nanoparticles (the best-performing catalyst) to optimize the active (edge and corner) sites that preferentially promote the C-C coupling reaction.

58 GEOSCIENCES↗

Energy-Efficient Capacitive Deionization through Electrode Modification and Process Development

Electrochemical separation technologies, such as capacitive deionization (CDI), are promising for addressing global energy and water challenges. However, there is a need to improve the performance, better understand property-performance relationships, and evaluate the longevity of CDI electrodes. This study explores the chemical modification of electrodes and the adjustment of CDI operating parameters. Results indicate that nitric acid (HNO3) conditioning of activated carbon cloth (ACC) electrodes removes metal oxides, introduces oxygen and nitrogen functionalities, and increases the specific capacitance (16% at 1 mV/s). Moreover, these changes in electrode properties positively impact device-level CDI performance. Through HNO3-conditioning of the ACC and tuning of the operational parameters, this work demonstrates higher electrosorption capacity (4.0x), greater charge efficiency (90% vs 24%), and lower energy consumption (3.8x). Despite these enhancements, limitations of the HNO 3 -conditioned ACC include decreased desorption kinetics and a 32% loss in electrosorption capacity after 200 cycles. Overall, this work provides guidance on using oxidative pretreatment via HNO 3 to modify ACC electrodes for CDI and evaluates the trade-offs associated with varying operational parameters.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Conserved Macromolecular Architecture of Poplar Secondary Cell Walls Revealed by ssNMR and Atomistic Modeling

The macromolecular architecture of plant secondary cell walls governs wood's mechanical and biochemical properties, yet its natural intra-species variability remains poorly characterized. Here, we combined 13C solid-state NMR (ssNMR), multivariate statistical analysis, and molecular modeling to profile nanoscale structure across 13 genetically diverse Populus trichocarpa genotypes grown in 13C-enriched atmospheres. SsNMR-derived phenotypes spanning composition, structure, mobility, and inter-polymer proximities reveal a conserved architecture, with a subtle yet coordinated variation organizing into dominant structural and secondary mobility axes. A representative atomistic model captures these features and reproduces experimental metrics. Molecular dynamics simulations support a weak but consistent positive correlation between cellulose abundance and crystalline-like order, with interior cellulose chains enriched in tg (trans-gauche) conformations without expanding crystalline cores. Together, experiment and simulation reveal a genetically buffered, broadly conserved nanoscale architecture across genotypes, where subtle fine-tuning of cellulose bundling and matrix packing balances mechanical performance with biological function.

09 BIOMASS FUELS↗

Neural architecture codesign for fast physics applications

We develop a pipeline to streamline neural architecture codesign for physics applications to reduce the need for ML expertise when designing models for novel tasks. Our method employs neural architecture search and network compression in a two-stage approach to discover hardware efficient models. This approach consists of a global search stage that explores a wide range of architectures while considering hardware constraints, followed by a local search stage that fine-tunes and compresses the most promising candidates. We exceed performance on various tasks and show further speedup through model compression techniques such as quantization-aware-training and neural network pruning. We synthesize the optimal models to high level synthesis code for FPGA deployment with the hls4ml library. Additionally, our hierarchical search space provides greater flexibility in optimization, which can easily extend to other tasks and domains. We demonstrate this with two case studies: Bragg peak finding in materials science and jet classification in high energy physics, achieving models with improved accuracy, smaller latencies, or reduced resource utilization relative to the baseline models.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Domain-specific text embedding model for accelerator physics

Accelerator physics presents unique challenges for natural language processing (NLP) due to its specialized terminology and complex concepts. A key component in overcoming these challenges is the development of robust text embedding models that transform textual data into dense vector representations, facilitating efficient information retrieval and semantic understanding. In this work, we introduce AccPhysBERT, a sentence embedding model fine-tuned specifically for accelerator physics. Our model demonstrates superior performance across a range of downstream NLP tasks, surpassing existing models in capturing the domain-specific nuances of the field. We further showcase its practical applications, including semantic paper-reviewer matching and integration into retrieval-augmented generation systems, highlighting its potential to enhance information retrieval and knowledge discovery in accelerator physics. Published by the American Physical Society 2025

Hellert, Thorsten (ORCID:0000000227970926)↗