Search NASA⌕ Search

SEARCH · Search NASA

Results for “Computational Efficiency”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 217 records · Page 12

Materials Learning Algorithms (MALA): Scalable machine learning for electronic structure calculations in large-scale atomistic simulations

We present the Materials Learning Algorithms (MALA) package, a scalable machine learning framework designed to accelerate density functional theory (DFT) calculations suitable for large-scale atomistic simulations. Using local descriptors of the atomic environment, MALA models efficiently predict key electronic observables, including local density of states, electronic density, density of states, and total energy. The package integrates data sampling, model training and scalable inference into a unified library, while ensuring compatibility with standard DFT and molecular dynamics codes. We demonstrate MALA's capabilities with examples including boron clusters, aluminum across its solid-liquid phase boundary, and predicting the electronic structure of a stacking fault in a large beryllium slab. Scaling analyses reveal MALA's computational efficiency and identify bottlenecks for future optimization. With its ability to model electronic structures at scales far beyond standard DFT, MALA is well suited for modeling complex material systems, making it a versatile tool for advanced materials research.

Density functional theory↗

Neural operator transformers capture bifurcating drift-wave turbulence in fusion plasma simulations

Self-consistent modeling of turbulence-driven transport is critical for optimizing confinement in magnetically confined fusion plasmas, such as tokamaks and stellarators. In particular, capturing the long-term co-evolution of turbulence, flow, and background plasma profiles remains computationally challenging. Direct numerical simulation of these multiscale, highly nonlinear processes is often demanding and impractical for real-time control or design optimization. To address this bottleneck, we investigate transformer-based neural operator partial differential equation surrogates for emulating the dynamics of drift-wave turbulence bifurcation mediated by zonal flows, using the modified Hasegawa–Wakatani (MHW) model as a prototypical system. We find that the finetuned neural operator model has excellent performance in capturing the multi-spatiotemporal-scales of MHW turbulence bifurcation and is robust to testing on rare and out-of-distribution dynamics. Specifically, we demonstrate that a single unified model accurately predicts both quasi-steady-state turbulence and a wide range of dynamical transition processes, such as nonlinear saturation, spontaneous suppression of turbulence, and the emergence of macroscopic zonal flows, over time horizons vastly exceeding the local turbulence correlation time. This computationally efficient approach establishes a strong foundation for fast, AI-based modeling of complex, multiscale phenomena in magnetized fusion plasmas.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Acceleration of Thermochemistry Solves in MOOSE and Pronghorn

This work focuses on the development and implementation of strategies to accelerate thermochemical calculations within MOOSE-based multiphysics simulations, particularly for applications in MSRs. We highlight the inherent complexity of nuclear materials, which require a multiscale approach to accurately model their behavior across various physical domains, including mechanical, chemical, and thermal phenomena. Thermochemical equilibrium calculations are crucial for predicting material properties and enhancing the fidelity of these simulations. The integration of Thermochimica, a Gibbs energy minimizer, into MOOSE allows for the direct minimization of Gibbs energy at every point on the mesh. However, the computational cost of such integration is significant. To address this, we explored acceleration strategies such as multi-threading support and the use of a thermodynamic ValueCache to reduce redundant calculations. Additionally, we investigated modifications to Thermochimica to enable phase constraints and improve its coupling with phase-field models, which are essential for simulating microstructural evolution and corrosion in MSR. These efforts aim to optimize the computational efficiency and accuracy of multiphysics simulations, thereby supporting the development of reliable and efficient nuclear materials for next-generation reactor technologies.

36 - MATERIALS SCIENCE↗

LDRD Abbreviated report: High-Order General-Discrete-Ordinates Method Enabling Efficient Deterministic Transport in Hydrodynamic Simulations

Deterministic transport simulations for national-security and energy applications often operate in high-dimensional phase-space, where accuracy and cost both become major challenges. A common numerical artifact in such problems is the “ray-effect,” which appears as unphysical streaks. Beyond misinterpretation, these artifacts can contaminate tightly coupled physics, such as fluid dynamics, radiation-hydrodynamics, and laser-plasma interactions, eroding the predictive capability of entire multiphysics workflows. Our objective was to make high-dimension studies practical on modern hardware while mitigating the ray-effect without relying on prohibitively expensive sampling approaches such as Monte Carlo methods. We developed the Generic Discretization Library (GenDiL), a Graphics Processing Unit (GPU)-first framework that uses high-order Discontinuous Galerkin (DG) methods and matrix-free algorithms to reduce memory usage and improve computational efficiency, critical for phase-space simulations. GenDiL supports phase-space adaptivity in both mesh size and polynomial order (hp-adaptivity) to place resolution only where it is needed. A central capability is Local Dimensional Refinement (LDR), which couples lower-dimension continuum models to higher-dimension kinetic models through stable and conservative interfaces, so that high-fidelity physics is applied only in regions where it is essential. Building on the GenDiL framework, we developed the General SN (GSN) family of algorithms as a true generalization of the polar SN approach (discrete ordinates, often denoted SN). Rather than tying discrete ordinates to a specific polar change of coordinates, GSN formulates transport on an arbitrary change of coordinates chosen to reduce ray-effect. We studied two complementary variants: an analytic variant, where the coordinate map is prescribed in advance by a closed-form function; and a data-driven variant, where a quantity of interest, such as the net flux, guides the coordinate system. GenDiL provides the library infrastructure for efficient GPU execution, but the GSN concept is algorithmic and independent of any one library. Across representative high-dimension tests, including non-symmetric solutions, both variants delivered strong ray-effect mitigation at practical cost, moving four- to six-dimensional analysis toward repeatable, routine studies.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

6D large charge and 2D Virasoro blocks

We compute observables in the interacting rank-one 6D 𝒩 =(2,0) superconformal field theory (SCFT) at large 𝑅-charge. We focus on correlators involving Φ 𝑛 , namely symmetric products of the bottom component of the supermultiplet containing the stress tensor. By using the moduli space effective action and methods from the large-charge expansion, we compute the operator product expansion coefficients ⟨Φ 𝑛 ⁢Φ 𝑚 ⁢Φ 𝑛+𝑚 ⟩ in an expansion in 1/𝑛. The coefficients of the expansion are only partially determined from the 6D perspective, but we manage to fix them order-by-order in 1/𝑛 numerically by utilizing the 6⁢D/2⁢D correspondence. This is made possible by the fact that this 6D observable can be extracted in 2D from a specific double-scaling limit of the vacuum Virasoro block, which can be efficiently computed numerically. We also extend the computation to higher-rank SCFTs, and discuss various applications of our results to 6D as well as 2D.

classical solutions in field theory↗

AutonomieAI: An efficient and deployable vehicle energy consumption estimation toolkit

Here, this paper presents AutonomieAI, a novel toolkit designed for efficient energy estimation of vehicles across diverse trip scenarios, routes, and drive cycles, applicable to a broad range of vehicle powertrain technologies. It leverages state-of-the-art Machine Learning techniques to deliver real-time energy prediction of vehicles, enabling co-simulation with transportation level system tools and opening doors for large-scale optimization at city, network or national level. Benchmark results show that AutonomieAI achieves high accuracy, with an average percentage error below 2% for most powertrain types, and computational efficiency capable of processing over 10,000 trips per second. Applications of AutonomieAI have potential to offer the flexibility to assist in solving eco-routing problems, optimize for vehicle and powertrain selection, study charging decision behavior, and optimize for charging station placement. AutonomieAI is the result of large neural network based model architectures, trained on very large and unique high fidelity vehicle simulation data. It is lightweight, deployable, efficient and has accuracy comparable to specialized and complex physics based simulation softwares.

Autonomie↗

Computing the Critical Temperature of the Affine-Transformed $D=3$ Ising Model Using Masked Autoregressive Flow

The simple Ising model provides a rich environment to build and study lattice field theories. As part of an ongoing project to construct a conformal field theory (CFT) on an arbitrarily curved manifold, in this work we develop methods to measure the critical temperature $β_c$ of the affine-transformed Ising model on the face-centered cubic (FCC) lattice. The main challenge in this endeavor is finding a computationally efficient and accurate method of interpolating and extrapolating Monte Carlo observables with respect to coupling coefficients and temperature. Herein, we compare two such methods. A traditional statistical approach uses the multiple histogram (MH) method, while a newer machine learning approach uses a masked autoregressive flow (MAF) to estimate the underlying probability density function of a set of observables. While the MH method is specifically designed to interpolate and extrapolate Monte Carlo observables, we find that MAF is a viable alternative for measuring $β_c$ with a computational cost that scales more favorably. Furthermore, we comment on additional advantages of MAF relevant to our work, such as extrapolating in system volume.

Svenson, Kai [Texas U.]↗

Accelerating LHC event generation with simplified pilot runs and fast PDFs

High-precision calculations are an indispensable ingredient for the success of the LHC physics programme, yet their poor computing efficiency has been a growing cause for concern, threatening to become a paralysing bottleneck in the coming years. We present solutions to eliminate the apprehension by focussing on two major components of generalpurpose Monte Carlo event generators: the evaluation of parton distribution functions, and the generation of perturbative matrix elements. We show that for the cost-driving event samples employed by the ATLAS experiment to model omnipresent, irreducible Standard Model backgrounds, such as weak boson or top-quark pair production in association with jets, these computational components dominate the overall run time by up to 80 %. We demonstrate that a reduction of the computing footprint of LHAPDF and SHERPA by factors of around 40 can be achieved for multi-leg NLO event generation.

Bothmann, Enrico [Gottingen U.]↗

Neuromorphic intermediate representation: A unified instruction set for interoperable brain-inspired computing

Abstract Spiking neural networks and neuromorphic hardware platforms that simulate neuronal dynamics are getting wide attention and are being applied to many relevant problems using Machine Learning. Despite a well-established mathematical foundation for neural dynamics, there exists numerous software and hardware solutions and stacks whose variability makes it difficult to reproduce findings. Here, we establish a common reference frame for computations in digital neuromorphic systems, titled Neuromorphic Intermediate Representation (NIR). NIR defines a set of computational and composable model primitives as hybrid systems combining continuous-time dynamics and discrete events. By abstracting away assumptions around discretization and hardware constraints, NIR faithfully captures the computational model, while bridging differences between the evaluated implementation and the underlying mathematical formalism. NIR supports an unprecedented number of neuromorphic systems, which we demonstrate by reproducing three spiking neural network models of different complexity across 7 neuromorphic simulators and 4 digital hardware platforms. NIR decouples the development of neuromorphic hardware and software, enabling interoperability between platforms and improving accessibility to multiple neuromorphic technologies. We believe that NIR is a key next step in brain-inspired hardware-software co-evolution, enabling research towards the implementation of energy efficient computational principles of nervous systems. NIR is available atneuroir.org

Science & Technology - Other Topics↗

LHC EFT WG note: SMEFT predictions, event reweighting, and simulation

This note provides a comprehensive overview of tools for predicting observables in the Standard Model effective field theory (SMEFT) at both tree level and one loop using event generators. We evaluate three primary methodologies–event reweighting, separate simulation of squared matrix elements, and full SMEFT process simulation–focusing on their statistical performance, computational efficiency, and potential biases. Each approach is assessed in terms of its accuracy, highlighting trade-offs between precision and resource demands. Practical insights into their applicability for high-energy physics analyses are offered, with particular attention to processes where SMEFT effects are significant. Additionally, we discuss the role of helicity in reweighting strategies and its impact on the quality of predictions. By comparing the methods across various LHC processes, this note provides guidance for selecting the most effective strategy for various SMEFT studies, ensuring robust predictions while optimizing computational resources.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Efficient sampling of free energy landscapes with functions in Sobolev spaces

Molecular simulations of biological and physical phenomena generally involve sampling complicated, rough energy landscapes characterized by multiple local minima. In this work, we introduce a new family of methods for advanced sampling that draw inspiration from functional representations used in machine learning and approximation theory. As shown here, such representations are particularly well suited for learning free energies using artificial neural networks. As a system evolves through phase space, the proposed methods gradually build a model for the free energy as a function of one or more collective variables, from both the frequency of visits to distinct states and generalized force estimates corresponding to such states. Implementation of the methods is relatively simple and, more importantly, for the representative examples considered in this work, they provide computational efficiency gains of up to several orders of magnitude over other widely used simulation techniques.

Approximation theory↗

Evaluating the Impact of Tritium Permeation Membrane Performance and Direct Internal Recycling on Fusion Fuel Cycle Efficiency Using TMAP8

An efficient fuel cycle is vital to sustainable and cost-effective energy generation in fusion systems. Since tritium is not widely available, fusion systems must breed their own tritium for sustainable fusion deuterium-tritium reactions. An inefficient fuel cycle increases the tritium inventory needed for operations, which increases costs, constraints on tritium management systems, and safety concerns. A fuel cycle model is a powerful tool for understanding tritium inventories and flow rates across all systems in the fuel cycle. By simplifying the technical details into time-dependent tritium flow rates and inventories, the model can simulate the entire fuel cycle with high computational efficiency, even for technologies that are still under development. It can therefore quantify the impact of new tritium management technologies on fuel cycle efficiency. To evaluate the impact of key components on reducing tritium inventory, we are using and expanding an existing fuel cycle models based on latest advancements in fuel cycle research. The new model integrates Tritium Permeation Membrane (TPM) and Direct Internal Recycling (DIR) to enhance tritium transport from blanket breeders and plasma exhaust. These fuel cycle models are implemented in TMAP8 (Tritium Migration Analysis Program, version 8), a MOOSE-based open-source application designed to provide cutting-edge capabilities for tritium transport and fuel cycle modeling. The study aims to demonstrate the extensibility of existing fuel cycle modeling capability in TMAP8 and to offer a proof-of-principle design for future fusion plant systems. The presentation will cover the performance of fuel cycle modeling capabilities available in TMAP8, highlight advancements in fuel cycle research, and present a sensitivity analysis of these models. The results underline potential approaches and technology solutions to lower tritium inventory requirements, highlighting their role in shaping the future of fusion energy.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Resolving local ordering and structure in Mn x Ge 1- x Te alloys through thermodynamic ensembles of pair distribution functions

Characterizing local bonding environments in complex materials is essential for understanding and optimizing their properties. Equally as important is the ability to predict local motifs as a function of synthesis conditions, enhancing chemists’ ability to design properties into materials. In this study, we present an approach to leverage statistical mechanics to generate temperature- and energy-informed ensemble averaged pair distribution functions (PDFs). This method, which we have named Thermodynamic Ensemble Averages of PDFs for Ordering and Transformations (TEAPOT), utilizes density functional theory (DFT) to relax supercells while incorporating energetic penalties for local order, enabling accurate and computationally efficient analysis of local structure. We apply this method to the neutron PDF measurements of the pseudobinary MnTe–GeTe (MGT) alloy, demonstrating its capability to resolve complex local distortions and chemical ordering. Our results reveal detailed insights into phase transformations and local distortions driven by Mn substitution. For compositions that globally present as rock salt, our analysis reveals that Ge coordination geometry is heavily impacted by synthesis temperature. We propose that high temperature synthesis conditions promote a lowered Ge polyhedra distortion, promoting high charge carrier mobility due to the alignment of local and global structure. Incorporating statistical mechanics and computation into experimental analysis thus guides synthesis of tailored local structure.

36 MATERIALS SCIENCE↗

TANTE: Time-adaptive operator learning via neural Taylor expansion

Operator learning for time-dependent partial differential equations (PDEs) has seen rapid progress in recent years, enabling efficient approximation of complex spatiotemporal dynamics. However, most existing methods rely on fixed time step sizes during rollout, which limits their ability to adapt to varying temporal complexity and often leads to error accumulation. In this work, we propose the Time-Adaptive Transformer with Neural Taylor Expansion (TANTE), a novel operator-learning framework that produces continuous-time predictions with adaptive step sizes. TANTE predicts future states by performing a Taylor expansion at the current state, where neural networks learn both the higher-order temporal derivatives and the local radius of convergence. This allows the model to dynamically adjust its rollout based on the local behavior of the solution, thereby reducing cumulative error and improving computational efficiency. We demonstrate the effectiveness of TANTE across a wide range of PDE benchmarks, achieving superior accuracy and adaptability compared to fixed-step baselines, delivering accuracy gains of 60-80 % and speed-ups of 30-40 % at inference time.

97 MATHEMATICS AND COMPUTING↗

Bias Correction and Statistical Downscaling of Future Solar Irradiance Projections Using the NSRDB

Assessing renewable energy resources under future climate scenarios has been highlighted to understand potential impacts of future climate change in renewable generation on the power sector. Climate model projection has been recognized by the renewable energy community as a useful data set to analyze the impacts of future climate change on renewable resources. However, future climate projections generated from general circulation models (GCMs) contain inherent biases that need to be corrected for accurate analysis of future projections of climate variables. In addition, the coarse spatiotemporal resolution of GCMs needs to be improved for regional climate studies. In this work, we develop statistical methods to downscale future projections of global horizontal irradiance (GHI) in a computationally efficient way. Our approach builds statistical downscaling models that correct bias of climate projection of GHI and downscale the future GHI projection from daily-scale to hourly-scale. The National Solar Radiation Database (NSRDB) is used to calibrate the statistical models and validate the downscaled GHI projections across the contiguous United State (CONUS). Preliminary results show that the statistical approach efficiently downscales climate projections of GHI with a nBIAS of 3%, nMAE of 34 % and nRMSE of 46% calculated against NSRDB for CONUS. This study describes the implemented methodology and initial results as well as future research to create high-resolution climate data sets for solar energy applications.

analytical models↗

A high-throughput framework for lattice dynamics

We develop an automated high-throughput workflow for calculating lattice dynamical properties from first principles including those dictated by anharmonicity. The pipeline automatically computes interatomic force constants (IFCs) up to 4th order from perturbed training supercells, and uses the IFCs to calculate lattice thermal conductivity, coefficient of thermal expansion, and vibrational free energy and entropy. It performs phonon renormalization for dynamically unstable compounds to obtain real effective phonon spectra at finite temperatures and calculates the associated free energy corrections. The methods and parameters are chosen to balance computational efficiency and result accuracy, assessed through convergence testing and comparisons with experimental measurements. Deployment of this workflow at a large scale would facilitate materials discovery efforts toward functionalities including thermoelectrics, contact materials, ferroelectrics, aerospace components, as well as general phase diagram construction.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Uncertainty Quantification of Fatigue Behavior of Rough AM Surfaces and Microstructures to Enable Hydrogen Gas Turbine

Modifying fossil-fueled industrial gas turbines to utilize low or zero-carbon fuels, such as hydrogen or hydrogen-natural gas blends, is a complex endeavor. The successful implementation of this technology hinges on three key design criteria: (1) developing new fuel injectors capable of efficiently burning alternative fuels, (2) ensuring manufacturability to meet cost and time-to-market goals, and (3) achieving component durability in the demanding environment of an operating gas turbine. Additive manufacturing (AM) accelerates product development, yet concerns persist regarding the durability of parts with rough AM surfaces. A fully experimental approach to quantify the fatigue performance of rough AM microstructures is both costly and labor-intensive. To address this, ORNL and Solar Turbines Incorporated (Solar) employed a crystal plasticity finite element (CPFE) model to identify the factors influencing AM surface fatigue behavior. These CPFE findings, combined with targeted experimental data, were used to develop a computationally efficient surrogate model suitable for assessing the lifespan of gas turbine engine components.

08 HYDROGEN↗

Optimization and Evaluation of Energy Savings for Connected and Autonomous Off-Road Vehicles

Off-road vehicles, such as wheel loaders, excavators, and harvesters, are extensively utilized across a wide range of industries, including construction, agriculture, and mining. These machines have become indispensable in supporting the day-to-day operational needs of a nation, playing a critical role in various sectors' infrastructure and productivity. However, despite their utility, off-road vehicles are significant consumers of fossil fuels, resulting in substantial emissions that contribute to environmental degradation. This highlights the pressing need for research and technological advancements aimed at improving their energy efficiency and reducing their carbon footprint. There are, however, two primary challenges that must be addressed to achieve these goals. First, off-road vehicles typically perform both driving and working tasks simultaneously, which introduces a high level of complexity into their overall dynamic systems. Analysis the interactions between these functions is challenging. Second, research into off-road vehicles is inherently interdisciplinary, demanding expertise across several domains such as fluid power systems, vehicle dynamics, control theory, optimization techniques, and real-world implementation. Recognizing these challenges, we proposed the project titled "Optimization and Evaluation of Energy Savings for Connected and Autonomous Off-Road Vehicles" as a comprehensive solution to enhance fuel efficiency while simultaneously improving productivity. This project specifically focuses on autonomous off-road vehicles, with particular attention to wheel loaders, and seeks to develop novel methods to optimize energy consumption without sacrificing operational performance. The project integrates real-time control algorithms, vehicle dynamics modeling, and co-optimization of powertrain system and vehicle system to achieve these goals. Our optimization strategy dynamically co-optimizes critical parameters at both the powertrain and vehicle levels, including vehicle speed, working tool movements, powertrain dynamics, and engine operations in real-time. To streamline this optimization process, we developed a vehicle model that captures the key dynamics while significantly enhancing computational efficiency. This allows the system to intelligently minimize fuel consumption, all while maintaining or even improving productivity through real-time calculations during various off-road operations. To validate the effectiveness of this energy optimization method, we introduced a state-of-the-art Hardware-in-the-Loop (HIL) testbed. This reconfigurable testbed seamlessly integrates the actual engine with virtual models of the wheel loader's subsystems, allowing for accurate emulation of real-world operational loads and environments. By simulating these conditions, the HIL testbed enables us to evaluate the wheel loader’s performance under diverse working scenarios, ensuring the developed solution is applicable in real-world operations. This testbed proved to be instrumental in validating the optimization algorithms and demonstrating the system's practical effectiveness. During the evaluation and testing phase, we employed the HIL testbed to rigorously assess the energy savings and productivity improvements generated by the optimized system. The results were highly encouraging, revealing that the automated wheel loader achieved over 30% fuel savings compared to traditional, human-operated cycles, with comparable or even enhanced levels of productivity. The insights gained from this HIL-based testing provided critical validation of our approach and highlighted the potential for deploying these optimized autonomous technologies in real-world off-road vehicles.

33 ADVANCED PROPULSION SYSTEMS↗