Search NASA⌕ Search

SEARCH · Search NASA

Results for “performance optimization”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 577 records · Page 32

Improving Additive Manufactured Component Performance through Multi-Scale Microstructure Simulation and Process Optimization

The purpose of this project was to utilize computational tools to understand the relationships between processing, microstructure, and properties for additively manufactured (AM) aluminum alloys for automotive applications, and to provide an engineering solution for helping to optimize process conditions. The project leverages ORNL developments in computational modeling, including AM process modeling, phase-field based microstructure evolution predictions, and data analytics techniques for mapping process conditions to material outcomes. The project utilized an Al-Cu-Mn-Zr alloy as a model material for studying formation of defects and microstructural features in response to variations in process conditions. Based on both pre-existing experimental data and simulation results, statistical process maps were constructed to identify regions of process space with minimal defect formation and advantageous microstructures and properties. The software tools used for this purpose were successful disseminated to GM, who were able to successful compile the relevant HPC codes within their own computing ecosystem and perform initial calculations to reproduce ORNL results.

36 MATERIALS SCIENCE↗

Surrogate Neural Architecture Codesign Package (SNAC-Pack)

Neural Architecture Search is a powerful approach for automating model design, but existing methods struggle to accurately optimize for real hardware performance, often relying on proxy metrics such as bit operations. We present Surrogate Neural Architecture Codesign Package (SNAC-Pack), an integrated framework that automates the discovery and optimization of neural networks focusing on FPGA deployment.

Weitz, Jason [UC, San Diego]↗

Surrogate Neural Architecture Codesign Package (SNAC-Pack)

Neural Architecture Search is a powerful approach for automating model design, but existing methods struggle to accurately optimize for real hardware performance, often relying on proxy metrics such as bit operations. We present Surrogate Neural Architecture Codesign Package (SNAC-Pack), an integrated framework that automates the discovery and optimization of neural networks focusing on FPGA deployment. SNAC-Pack combines Neural Architecture Codesign's multi-stage search capabilities with the Resource Utilization and Latency Estimator, enabling multi-objective optimization across accuracy, FPGA resource utilization, and latency without requiring time-intensive synthesis for each candidate model. We demonstrate SNAC-Pack on a high energy physics jet classification task, achieving 63.84% accuracy with resource estimation. When synthesized on a Xilinx Virtex UltraScale+ VU13P FPGA, the SNAC-Pack model matches baseline accuracy while maintaining comparable resource utilization to models optimized using traditional BOPs metrics. This work demonstrates the potential of hardware-aware neural architecture search for resource-constrained deployments and provides an open-source framework for automating the design of efficient FPGA-accelerated models.

Weitz, Jason [UC, San Diego] (ORCID:00090004631535↗

Dynamic Modeling of a Kaplan Hydroturbine Using Optimal Parametric Tuning and Real Plant Operational Data

To address grid variability caused by renewable energy integration and to maintain grid reliability and resilience, hydropower must quickly adjust its power generation over short time periods. This changing energy generation landscape requires advance technology integration and adaptive parameter optimization for hydropower systems via digital twin effort. However, this is difficult owing to the lack of characterization and modeling for the nonlinear nature of hydroturbines. To solve this issue, this paper first formulates a six-coefficient Kaplan hydroturbine model and then proposes a parametric optimization tuning framework based on the Nelder–Mead algorithm for adaptive dynamic learning of the six-coefficients so as to build models that describe the turbine. To assess the performance of the proposed optimal parametric tuning technique, operational data from a real-world Kaplan hydroturbine unit are collected and used to model the relationship between the gate opening and the generated power production. The findings show that the proposed technique can effectively and adaptively learn the unknown dynamics of the Kaplan hydroturbine while optimally tune the unknown coefficients to match the generated power output from the real hydroturbine unit with an inaccuracy of less than 5%. The method can be used to provides optimal tuning of parameters critical for controller design, operational optimization and daily maintenance for hydroturbines in general.

13 HYDRO ENERGY↗

Proliferated Resilient Economical Half-Meter Aperture Space Telescopes (PREEMPT)

The PREEMPT LDRD was motivated by a major national need for better space-based imaging systems that are both high performing and affordable for the US Government. Current optical payloads for intelligence, surveillance and reconnaissance (ISR) and space domain awareness (SDA) missions can cost hundreds of millions of dollars and take years to develop, which makes it difficult to build the large constellations required for persistent, world-wide coverage. To address this, the team aimed to advance a different kind of large-aperture (>25 cm) telescope, called a monolithic telescope, in which key optical surfaces are built into a single piece of fused silica. This design greatly reduces payload and spacecraft complexity, the need for precision focus actuators, improves mechanical and thermal robustness, and lowers cost when compared with traditional Cassegrain telescopes that rely on many precisely aligned components. The project focused on five primary thrusts. The first thrust was to advance the concept of a V10, 25 cm, monolithic telescope forward from optical design to flight-ready stage. This was accomplished in partnership with Optimax, who delivered the first test unit in the early stages of the LDRD. The team developed several technologies necessary for this optic to be integrated into a flight demonstration. These include carbon fiber housings, highly detailed structural and thermal models and stress-reducing elastic averaging Hirth groove designs. These technologies resulted in a successful maturation of the optic, which is now slated to fly in late 2026/early 2027 for a demonstration mission. The second thrust was to advance the manufacturability of these optics. In collaboration with NIF’s optical manufacturing shop, we reduced polishing time from 480 hours to 65 hours through the implementation of optimized processes and new tools. The NIF team utilized a conceptual V8 (18 cm) optic to demonstrate this optimization, though it can be applied to the rest of the monolithic optic portfolio. Third, the team developed the first conceptual V20 (50 cm) payload, which is slated to be the next generation of LLNL optical payload systems. A set of structural, dynamic and thermal simulations were performed to identify potential challenges in the future development of this payload. Early-stage simulations suggest the payload is feasible, though thermal management will be the key focus area to maintain optimal performance. Fourth, a non-linear model of Viton was developed, to further enhance the reliability of our structural and dynamic models for future payloads. Viton acts as the primary interface material between the optic and its housing. Lastly, the team focused on successfully displaying the feasibility of using additively manufactured metal composites for optical space payloads. The team successfully demonstrated layer by layer deposition of Al-SiC composites, which have highly tunable structural and coefficient of thermal expansion (CTE) properties. These are crucial for optical payloads because CTE mismatch is one of the causes for degraded optical performance for telescopes in orbit. Overall, the work showed that monolithic telescopes could become a practical, lower-cost path to high-resolution space imaging for both national security and scientific missions.

42 ENGINEERING↗

Distributionally Robust Variational Quantum Algorithms With Shifted Noise

Given their potential to demonstrate near-term quantum advantage, variational quantum algorithms (VQAs) have been extensively studied. Although numerous techniques have been developed for VQA parameter optimization, it remains a significant challenge. A practical issue is the high sensitivity of quantum noise to environmental changes, and its propensity to shift in real time. This presents a critical problem as an optimized VQA ansatz may not perform effectively under a different noise environment. For the first time, we explore how to optimize VQA parameters to be robust against unknown shifted noise. We model the noise level as a random variable with an unknown probability density function (PDF), and we assume that the PDF may shift within an uncertainty set. This assumption guides us to formulate a distributionally robust optimization problem, with the goal of finding parameters that maintain effectiveness under shifted noise. We utilize a distributionally robust Bayesian optimization solver for our proposed formulation. This provides numerical evidence in both the Quantum Approximate Optimization Algorithm (QAOA) and the Variational Quantum Eigensolver (VQE) with hardware-efficient ansatz, indicating that we can identify parameters that perform more robustly under shifted noise. We regard this work as the first step towards improving the reliability of VQAs influenced by real-time noise.

97 MATHEMATICS AND COMPUTING↗

Targeted Adaptive Design

Modern advanced manufacturing and advanced materials design often require searches of relatively high-dimensional process control parameter spaces for settings that result in optimal structure, property, and performance parameters. The mapping from the former to the latter must be determined from noisy experiments or from expensive simulations. Here, we abstract this problem to a mathematical framework in which an unknown function from a control space to a design space must be ascertained by means of expensive noisy measurements, which locate control settings generating desired design features within specified tolerances, with quantified uncertainty. We describe targeted adaptive design (TAD), a new algorithm that performs this sampling task efficiently. TAD creates a Gaussian process surrogate model of the unknown mapping at each iterative stage, proposing a new batch of control settings to sample experimentally and optimizing the updated expected log-predictive probability density of the target design. TAD either stops upon locating a solution with uncertainties that fit inside the tolerance box or uses a measure of expected future information to determine that the search space has been exhausted with no solution. TAD thus embodies the exploration-exploitation tension in a manner that recalls, but is essentially different from, Bayesian optimization and optimal experimental design.

97 MATHEMATICS AND COMPUTING↗

Development of Industrial Scale Rare Earth Master Alloys from Their Native Oxides for Magnet Production

The objective of the project is to develop an energy-efficient, reduced cost, and single-step critical metal oxide reduction and alloying methodology for the production of NdFeB and SmCo magnets to facilitate the establishment of a sustainable domestic critical materials supply chain. The method consists of an immiscible molten salt flux layer and a higher density molten metal alloy pool (FeB or Co). The rare earth (RE) oxide and a reductant are added to the molten salt layer, where the reductant first strips the oxygen from the rare earth oxide. Subsequently, the separated RE metal diffuses into the molten metal pool below, creating a RE-saturated master alloy. Towards this end, thermodynamic calculations of the reactions between RE oxides, metallic reducing agent, and molten salt bath chemistry have been performed, and the ideal feeds and conditions for extraction and diffusion to produce master alloys were established. Small-scale and scaled experimentation was performed to validate and optimize the feasibility of viable reactions, temperatures, and process conditions with regards to yield, composition, and process efficiency. Finally, full-scale experiments for the production of NdFeB and SmCo magnets were performed, and their performance characteristics were established.

36 MATERIALS SCIENCE↗

Influence of the as-built microstructure on the recrystallization of an additively manufactured Inconel939 Ni-based superalloy

This study investigates the influence of the as-built microstructure on the recrystallization (RX) behavior and mechanical properties of the Ni-based superalloy Inconel 939 produced by laser powder bed fusion (PBF-LB/M). Two distinct as-built microstructures were obtained by varying the hatch distance (h d ): a columnar, strongly textured condition (h d =50, termed h d 50) and an equiaxed, weakly textured condition (h d =70, termed h d 70)). Both were subjected to nine solution treatments combining three temperatures (1100, 1150, and 1200 °C) and three holding times (1, 4, and 8 h). Comprehensive microstructural characterization was conducted to assess grain morphology, texture, grain boundary character, dislocation density, and precipitate distribution. Recrystallization was found to be significantly slower than in cast counterparts, requiring higher temperatures and longer times for completion. The initial microstructure plays a decisive role: full RX was achieved only in hd70 specimens after treatment at 1200 °C for 8 h, whereas hd50 samples exhibited delayed and incomplete RX under identical conditions. This behavior is attributed to the finer grain size and higher fraction of high-angle grain boundaries in hd70, which promote recrystallization. Mechanical testing revealed that hd70 samples subjected to a 1200 °C/8 h treatment followed by standard double ageing show higher yield and tensile strengths across the investigated temperature range than both printed and cast Inconel939 processed under conventional conditions, albeit with slightly reduced ductility. The enhanced mechanical performance is attributed to the larger grain size, which limits grain boundary sliding. These results demonstrate the critical importance of controlling the as-built microstructure and tailoring post-processing strategies to optimize high-temperature performance of PBF-LB/M Inconel939.

Inconel939↗

ZEUS: An Efficient GPU Optimization Method Integrating PSO, BFGS, and Automatic Differentiation

We introduce a novel, efficient computational method, ZEUS, for numerical optimization, and provide an open-source implementation. It has four key ingredients: (1) particle swarm optimization (PSO), (2) the use of the Broyden-Fletcher-Goldfarb-Shanno (BFGS) method, (3) automatic differentiation (AD), and (4) GPUs. Our approach addresses the computational challenges inherent in high-dimensional, non-convex optimization problems. In the first phase of the algorithm, we get a potentially good set of starting points using PSO. Thereafter, we run BFGS independently in parallel from these starting points. BFGS is one of the best-performing algorithms for numerical optimization. However, it requires the gradient of the function being optimized. ZEUS integrates automatic differentiation into BFGS thus avoiding the need for the user to calculate derivatives explicitly. The use of GPUs allows ZEUS to speed up the calculations substantially. We carry out systematic studies to explore the trade-offs between the number of PSO iterations taken, starting points, and BFGS iteration depth. We show that a handful of iterations of PSO can improve global convergence when combined with BFGS. We also present performance studies using common test functions. The source code can be found at https://github.com/fnal-numerics/global-optimizer-gpu.

Soos, Dominik [Old Dominion U.]↗

Superlative mechanical energy absorbing efficiency discovered through self-driving lab-human partnership

Energy absorbing efficiency is a key determinant of a structure’s ability to provide mechanical protection and is defined by the amount of energy that can be absorbed prior to stresses increasing to a level that damages the system to be protected. Here, we explore the energy absorbing efficiency of additively manufactured polymer structures by using a self-driving lab (SDL) to perform >25,000 physical experiments on generalized cylindrical shells. We use a human-SDL collaborative approach where experiments are selected from over trillions of candidates in an 11-dimensional parameter space using Bayesian optimization and then automatically performed while the human team monitors progress to periodically modify aspects of the system. The result of this human-SDL campaign is the discovery of a structure with a 75.2% energy absorbing efficiency and a library of experimental data that reveals transferable principles for designing tough structures.

42 ENGINEERING↗

Selective capture and recovery of uranium oxide colloids from aqueous soil suspensions using high gradient magnetic filtration

High Gradient Magnetic Filtration (HGMF) is a promising method for the selective capture and recovery of uranium oxide from surface soils. To date, however, magnetic filtration of uranium oxide has only been demonstrated at a proof-of-principle scale using relatively small filters (<5 cm 3 ) at low flowrates (<60 mL/min). Here, to explore the efficacy of magnetic filtration of uranium oxide at a larger scale, a newly designed HGMF apparatus that is more than an order of magnitude larger than our earlier filters (106 cm 3 ) was designed, fabricated, and tested at relatively high flowrates. Filtration experiments were performed using aqueous uranium oxide particle suspensions with Arizona Road Dust (ARD) as a soil simulant. At a flowrate of 125 mL/min, the apparatus’ uranium capture rate was exceptionally high (96 %), but selectivity was poor due to the high rate of capture for diamagnetic soil constituents (e.g., 77 % for silicon). All particles were captured at a lower rate when the flowrate was increased to 250 mL/min, but uranium selectivity was significantly increased due to the more substantial reduction in diamagnetic particle capture (i.e., capture rate of 77 % and 15 % for uranium and silicon, respectively). When backwashing the apparatus at the same flowrates used during filtration experiments, the rate of uranium recovery tended to be fairly low. Nevertheless, higher flowrates (1 L/min) and sonication were both shown to be highly effective methods of increasing uranium recovery. Magnetic field simulations were also performed to investigate potential optimizations to the design of the apparatus. These simulations showed that the intensity of the applied magnetic field could be increased by increasing the thickness of the steel magnetic housing. Additionally, stochastic trajectory simulations were performed to investigate the potential mechanisms of particle capture.

HGMF↗

Assessing the Performance and Impact of PV Technologies on Storage in Hybrid Renewable Systems

Traditional monofacial photovoltaic (mPV) systems are commonly adopted and well-documented because of their lower upfront costs in comparison to bifacial photovoltaic (bPV) systems. This study investigates how PV technologies impact energy storage in grid-scale hybrid renewable systems, focusing on optimizing and assessing the performance of mPV and bPV technologies integrated with pumped storage hydropower. Using Ludington City, Michigan as a case study and analyzing real-world data such as solar irradiance, ambient temperature, and utility-scale load profiles, the research highlights the operational and economic benefits of bPV systems. The results reveal that bPV systems can pump approximately 10.38% more water annually to the upper reservoir while achieving a lower levelized cost of energy ($0.0578/kWh for bPV vs. $0.0672/kWh for mPV). This study underscores the outstanding potential of bPV systems in enhancing energy storage and management strategies, contributing to a more sustainable and resilient renewable energy future.

13 HYDRO ENERGY↗

Rapid neutron and gamma-ray source localization using machine learning

Rapid localization of radiation sources is critical for applications including nuclear emergency response, safeguards, and security. However, conventional imaging systems such as neutron scatter cameras and Compton cameras depend on rare coincidence events, which often result in long acquisition times. In this work, we address the challenge of rapid source localization by developing a machine learning approach to predict the direction of a single radiation source using only count rates from an array of neutron and gamma-ray detectors. The proposed model is a fully connected neural network (FCNN) trained using Monte Carlo simulation data from a 252 Cf source. The model hyperparameters are optimized with a small set of routine 252 Cf measurements. We benchmarked the performance of the trained and optimized machine learning model using additional 252 Cf , 137 Cs , and PuBe measurements under laboratory conditions with varying source-detector configurations. For these measurements, the machine learning model achieved a mean localization error smaller than 30° with 3 x 10 3 system counts, corresponding to 8 s measurement time for the imaging system used in this work. In this low-statistics regime, the method outperformed traditional scatter-based imaging by more than 75% in localization accuracy for the evaluated measurement configurations. These results demonstrate that a machine learning-based approach can significantly reduce the time required for accurate single-source localization, providing a robust and computationally efficient alternative to traditional imaging systems in time-critical nuclear security and emergency response scenarios.

Gamma-ray imaging↗

Enabling Efficient Sparse Computations using Linear Algebra Aware Compilers

This project developed the LAPIS compiler framework, built on the Multilevel Intermediate Representation (MLIR), to optimize sparse linear algebra operations and support performance portability across diverse architectures. The main innovation of LAPIS is the Kokkos dialect, which allows for lowering codes from a high productivity language to different architectures in an elegant way. The dialect also allows the conversion of lower-level MLIR code to C++ Kokkos code, facilitating the integration of scientific machine learning (SciML) models into applications. To extend LAPIS for distributed memory architectures, a new partition dialect was created to manage the distribution of sparse tensors and express communication patterns for sparse linear algebra operations. This dialect also supports the distributed execution of operators and includes algorithmic optimizations to minimize communication to improve performance. The project also demonstrates that MLIR can enable effective linear algebra-level optimizations, improving performance on different GPUs for both sparse and dense linear algebra kernels. Key applications of LAPIS include sparse linear algebra and graph kernels, TenSQL, a relational database management solution built on GraphBLAS, and the development of subgraph isomorphism and monomorphism kernels, showcasing performance portability. In summary, the LAPIS framework supports productivity, performance, portability, and distributed memory execution, while also enabling linear algebra-level optimizations that are challenging in traditional programming languages, with successful applications ranging from simple sparse linear algebra to complex graph kernels.

97 MATHEMATICS AND COMPUTING↗

Empirical thermophotovoltaic performance predictions and limits

Significant progress has been made in the field of thermophotovoltaics, with efficiency recently rising to over 40% due to improvements in cell design and material quality, higher emitter temperatures, and better spectral management. However, inconsistencies in trends for efficiency with semiconductor bandgap energy across various temperatures pose challenges in predicting optimal bandgaps or expected performance for different applications. To address these issues, here we present realistic performance predictions for various types of single-junction cells over a broad range of emitter temperatures using an empirical model based on past cell measurements. Our model is validated using data from different authors with various bandgaps and emitter temperatures, and an excellent agreement is seen between the model and the experimental data. Using our model, we show that in addition to spectral losses, it is important to consider practical electrical losses associated with series resistance and cell quality to avoid overestimation of system efficiency. Here, we also show the effect of modifying various system parameters such as bandgap, above and below-bandgap reflectance, saturation current, and series resistance on the efficiency and power density of thermophotovoltaics at different temperatures. Finally, we predict the bandgap energies for best performance over a range of emitter temperatures for different cell material qualities.

14 SOLAR ENERGY↗

Quantification of trace iodine using laser-induced breakdown spectroscopy for real-time monitoring of nuclear off-gas streams

This study evaluated the potential of laser-induced breakdown spectroscopy (LIBS) for real-time monitoring of trace gas-phase iodine, which is an element of high significance in nuclear applications due to its long radioactive half-life (as iodine-129), volatility, and biological impact. In anticipation of iodine evolving into off-gas systems in molten salt reactor and nuclear fuel recycling applications, this research aimed to assess LIBS performance in flowing argon and helium matrices; optimize measurement parameters using a multichannel spectrometer; and perform calibrations to assess predictive capabilities and limits of detection (LODs). Experimental results successfully measured gas-phase iodine in flowing argon and helium; however, trace iodine was not detected in air. Optimal delay times were determined to be 10 µs for argon and 1 µs for helium, which are consistent with the expected shorter plasma lifetime in helium relative to argon. An emission line survey was provided with the 206.16, 804.37, 902.24, and 905.83 nm peaks, which were identified as the strongest emission peaks. Calibration models were successfully built in both helium and argon, achieving LODs down to 3 ppm in helium and 5 ppm in argon. The iodine emission at 905.83 nm emerged as the most robust for calibration and was subsequently applied to a time series dataset in argon. The predictive trace confirmed the feasibility of employing LIBS for continuous, online quantification of trace iodine in flowing gas systems.

Andrews, Hunter B. [Oak Ridge National Laboratory ↗