Search NASA⌕ Search

SEARCH · Search NASA

Results for “Energy Efficient Computing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 217 records · Page 12

Understanding Reliability Trade-Offs in 1T-nC and 2T-nC FeRAM Designs

Ferroelectric random access memory (FeRAM) is a promising candidate for energy-efficient nonvolatile memory, particularly for logic-in-memory and compute-in-memory (CIM) applications. Among the available cell architectures, One-Transistor–n-Capacitor (1T-nC) and two-transistor–n-capacitor (2T-nC) FeRAMs each offer distinct trade-offs in density, scalability, and reliability. In this work, we present a comparative study of these two architectures under both dimensional scaling ( XY/Z shrinkage) and vertical integration (increasing stacked capacitors per cell). Using technology computer-aided design (TCAD) and circuit-level simulations, we analyze how scaling impacts ferroelectric capacitance, parasitic coupling, and floating-node (FN) dynamics, which together dictate sense margin (SM) and read stability. A key mitigation strategy—floating unselected capacitors—is applied to both architectures, effectively decoupling the SM from the number of stacked capacitors and enabling tractable analysis across scaling regimes. Results show that 1T-nC suffers more from charge sharing with the bitline (BL), while 2T-nC benefits from transistor isolation and stronger low-voltage sensing at the cost of increased area. By systematically evaluating these behaviors across scaling directions, this work establishes the reliability trade-offs of 1T-nC and 2T-nC cells and provides design guidelines for high-density, vertically integrated FeRAM systems.

1T-nC↗

High-performance finite elements with MFEM

The MFEM (Modular Finite Element Methods) library is a high-performance C++ library for finite element discretizations. MFEM supports numerous types of finite element methods and is the discretization engine powering many computational physics and engineering applications across a number of domains. Furthermore, this paper describes some of the recent research and development in MFEM, focusing on performance portability across leadership-class supercomputing facilities, including exascale supercomputers, as well as new capabilities and functionality, enabling a wider range of applications. Much of this work was undertaken as part of the Department of Energy’s Exascale Computing Project (ECP) in collaboration with the Center for Efficient Exascale Discretizations (CEED).

97 MATHEMATICS AND COMPUTING↗

Integrating ytopt and libEnsemble to autotune OpenMC

Ytopt is a Python machine-learning-based autotuning software package developed within the ECP PROTEAS-TUNE project. The ytopt software adopts an asynchronous search framework that consists of sampling a small number of input parameter configurations and progressively fitting a surrogate model over the input-output space until exhausting the user-defined maximum number of evaluations or the wall-clock time. libEnsemble is a Python toolkit for coordinating workflows of asynchronous and dynamic ensembles of calculations across massively parallel resources developed within the ECP PETSc/TAO project. libEnsemble helps users take advantage of massively parallel resources to solve design, decision, and inference problems and expands the class of problems that can benefit from increased parallelism. In this paper we present our methodology and framework to integrate ytopt and libEnsemble to take advantage of massively parallel resources to accelerate the autotuning process. Specifically, we focus on using the proposed framework to autotune the ECP ExaSMR application OpenMC, an open source Monte Carlo particle transport code. OpenMC has seven tunable parameters some of which have large ranges such as the number of particles in-flight, which is in the range of 100,000 to 8 million, with its default setting of 1 million. Setting the proper combination of these parameter values to achieve the best performance is extremely time-consuming. Therefore, we apply the proposed framework to autotune the MPI/OpenMP offload version of OpenMC based on a user-defined metric such as the figure of merit (FoM) (particles/s) or energy efficiency energy-delay product (EDP) on Crusher at Oak Ridge Leadership Computing Facility. In conclusion, the experimental results show that we achieve the improvement up to 29.49% in FoM and up to 30.44% in EDP.

Autotuning↗

CSP: A Multifaceted Hybrid Architecture for Space Computing

Research on the CHREC Space Processor (CSP) takes a multifaceted hybrid approach to embedded space computing. Working closely with the NASA Goddard SpaceCube team, researchers at the National Science Foundation (NSF) Center for High-Performance Reconfigurable Computing (CHREC) at the University of Florida and Brigham Young University are developing hybrid space computers that feature an innovative combination of three technologies: commercial-off-the-shelf (COTS) devices, radiation-hardened (RadHard) devices, and fault-tolerant computing. Modern COTS processors provide the utmost in performance and energy-efficiency but are susceptible to ionizing radiation in space, whereas RadHard processors are virtually immune to this radiation but are more expensive, larger, less energy-efficient, and generations behind in speed and functionality. By featuring COTS devices to perform the critical data processing, supported by simpler RadHard devices that monitor and manage the COTS devices, and augmented with novel uses of fault-tolerant hardware, software, information, and networking within and between COTS devices, the resulting system can maximize performance and reliability while minimizing energy consumption and cost. NASA Goddard has adopted the CSP concept and technology with plans underway to feature flight-ready CSP boards on two upcoming space missions.

Reconfigurable↗

Computationally efficient method for determining limiting velocities of edge dislocations in anisotropic crystals

The continuum-limit theory of dislocations in crystals predicts divergences in the elastic energy at crystal-geometry dependent limiting velocities vL, which separate subsonic, transsonic, and supersonic dislocation glide regimes and are therefore import for material strength models at high strain rates. Although it is known how to calculate those limiting velocities, there is one special case - edge dislocations with reflection symmetry, but non-vanishing elastic constants c16 or c26 - where previous methods have been notoriously slow. In this letter, we address this deficiency by deriving a computationally efficient method for determining the limiting velocities of edge dislocations with reflection symmetry which is two orders of magnitude faster than the previous method.

36 MATERIALS SCIENCE↗

Educational Consortium for Energy-related Data Science & Computation in Building Engineering Programs

The project spearheaded by Pennsylvania State University aims to address the growing need for integrating energy-focused computation and data science into building engineering education. As the demand for energy-efficient building designs and operations increases, the educational sector must adapt to equip future engineers with the necessary skills. This initiative responds to this need by developing a consortium that unites multiple institutions to enhance curriculum development, dataset curation, and resource sharing, thereby ensuring students are well-prepared for the evolving energy sector. The primary goal of the project is to establish a consortium that will develop and disseminate educational materials and training programs focused on energy-related data science and computation. Key accomplishments include the creation of a beta website for resource sharing, the development of training programs and standalone modules, and the curation of datasets accessible to the public. This effort will culminate in a curriculum that incorporates advanced modeling technologies and data science skills into building engineering programs.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Understanding Adsorption and Reactions at Aqueous Oxide Interfaces with Neural Network Potential Molecular Dynamics

Chemical processes at metal oxide−water interfaces are of central importance in geochemistry, biology, and energy technologies. A better understanding of these processes would allow us to make a significant step toward optimizing and controlling them, which could in turn lead to broader impacts. Computational modeling is indispensable to accomplishing this task because complexity and disorder often make it difficult to extract atomistic information from experiments. Balancing computational cost and accuracy, simulation schemes based on efficient machine learning representations of the potential energy surface (PES) predicted by ab initio calculations have become increasingly popular over the past decade. In particular, several studies have demonstrated the ability of machine learning models to accurately reproduce the complex ab initio PESs of aqueous oxide interfaces, allowing simulations of systems and processes that are not accessible with ab initio methods. In this Account, we review our recent efforts to understand adsorption processes and reactions at aqueous oxide interfaces using deep potential molecular dynamics (DPMD), a simulation scheme employing deep neural networks (DNNs), which has proven to be quite successful in accurately describing many different systems in the condensed phase. After summarizing the DPMD methodology, we first review our work on the acid−base chemistry of oxide surfaces in contact with water, a fundamental characteristic that controls proton transfer and surface charge at the interface. We focus on the aqueous interface of rutile IrO 2 , an oxide material thus far considered the best catalyst for the oxygen evolution reaction (OER). We show that this interface is characterized by a large fraction of dissociated water and a strong Brønsted acidity of the surface sites, in good agreement with the experimentally measured value of the point of zero proton charge. In our second example, we investigate how the adsorption of organic species from ambient air or water affects the structure and wettability of the aqueous interfaces of TiO 2 , a prototypical photocatalytic material. This is a question that is relevant to understanding the UV-induced hydrophilicity of TiO 2 surfaces, a property at the basis of self-cleaning windows and related applications. Specifically focusing on formic and acetic acids, the two most common atmospheric organic acids, our simulations reveal that these acids control the wettability of TiO 2 largely through acid−base chemistry at the interface rather than chemisorption on the oxide surface, a finding that could help improve the design of self-cleaning surfaces and photocatalytic devices. Finally, we review our recent study of methanol at TiO 2 −water interfaces, a system whose interest is largely motivated by the role of methanol in enhancing photocatalytic hydrogen evolution on TiO 2 . Our simulations provide mechanistic insights into the coupled roles of the organic adsorbate and water at the TiO 2 interface, with implications for how methanol enhances the activity of H 2 evolution.

adsorption↗

Averaging techniques for steady and unsteady calculations of a transonic fan stage

It is often desirable to characterize a turbomachinery flow field with a few lumped parameters such as total pressure ratio or stage efficiency. Various averaging schemes may be used to compute these parameters. Here three averaging schemes, the momentum, energy, and area averaging schemes, are described and compared for two computed solutions of the midspan section of a transonic fan stage: a steady averaging-plane solution in which average rotor outflow conditions were used as stator inflow conditions and an unsteady rotor-stator interaction solution. The unsteady solution is described, some unsteady flow phenomena are discussed and the steady pressure distributions are compared. Despite large unsteady pressure fluctuations on the stator surface, the steady pressure distribution matched the average unsteady distribution almost exactly. Stator wake profiles, stator loss coefficient, and stage efficiency were computed for the two solutions with the three averaging schemes and are compared. In general the energy averaging scheme gave good agreement between the averaging-plane solution and the time-averaged unsteady solution, even though certain phenomena due to unsteady wake migration were neglected.

Wyss, M. L.↗

Theoretical investigation of wave-vector-dependent analytical and numerical formulations of the interband impact-ionization transition rate for electrons in bulk silicon and GaAs

The electron interband impact-ionization rate for both silicon and gallium arsenide is calculated using an ensemble Monte Carlo simulation with the expressed purpose of comparing different formulations of the interband ionization transition rate. Specifically, three different treatments of the transition rate are examined: the traditional Keldysh formula, a new k-dependent analytical formulation first derived by W. Quade, E. Scholl, and M. Rudan (1993), and a more exact, numerical method of Y. Wang and K. F. Brennan (1994). Although the completely numerical formulation contains no adjustable parameters and as such provides a very reliable result, it is highly computationally intensive. Alternatively, the Keldysh formular, although inherently simple and computationally efficient, fails to include the k dependence as well as the details of the energy band structure. The k-dependent analytical formulation of Quade and co-workers overcomes the limitations of both of these models but at the expense of some new parameterization. It is found that the k-dependent analytical method of Quade and co-workers produces very similar results to those obtained with the completely numerical model for some quantities. Specifically, both models predict that the effective threshold for impact ionization in GaAs and silicon is quite soft, that the majority of ionization events originate from the second conduction band in both materials, and that the transition rate is k dependent. Therefore, it is concluded that the k-dependent analytical model can qualitatively reproduce results similar to those obtained with the numerical model yet with far greater computational efficiency. Nevertheless, there exist some important drawbacks to the k-dependent analytical model of Quade and co-workers: These are that it does not accurately reproduce the quantum yield data for bulk silicon, it requires determination of a new parameter, related physically to the overlap intergrals of the Bloch state which can only be adjusted by comparison to experiment, and fails to account for any wave-vector dependence of the overlap integrals. As such the transition rate may be overestimated at those points for which 'near vertical,' small change in k, transitions occur.

Kolnik, Jan↗

Theoretical Investigation of Wave-Vector-Dependent Analytical and Numerical Formulations of the Interband Impact-Ionization Transition Rate for Electron in Bulk Silicon and GaAs

The electron interband impact-ionization rate for both silicon and gallium arsenide is calculated using an ensemble Monte Carlo simulation with the expressed purpose of comparing different formulations of the interband ionization transition rate. Specifically, three different treatments of the transition rate are examined: the traditional Keldysh formula, a new k-dependent analytical formulation first derived by W. Quade, E Scholl, and M. Rudan, and a more exact, numerical method of Y. Wang and K. F. Brennan. Although the completely numerical formulation contains no adjustable parameters and as such provides a very reliable result, it is highly computationally intensive. Alternatively, the Keldysh formula, although inherently simple and computationally efficient, fails to include the k dependence as well as the details of the energy band structure. The k-dependent analytical formulation of Quade and co-workers overcomes the limitations of both of these models but at the expense of some new parameterization. It is found that the k-dependent analytical method of Quade and co-workers produces very similar results to those obtained with (he completely numerical model for some quantities. Specifically, both models predict that the effective threshold for impact ionization in GaAs and silicon is quite soft, that the majority of ionization events originate from the second conduction band in both materials, and that the transition rate is k dependent. Therefore, it is concluded that the k-dependent analytical model can qualitatively reproduce results similar to those obtained with the numerical model yet with far greater computational efficiency. Nevertheless, there exist some important drawbacks to the k-dependent analytical model of Quade and co-workers: These are that it does not accurately reproduce the quantum yield data for bulk silicon, it requires determination of a new parameter, related physically to (he overlap integrals of the Bloch state which can only be adjusted by comparison to experiment, and fails to account for any wave-vector dependence of the overlap integrals. As such [he transition rate may be overestimated at those points for which "near vertical," small change in k, transitions occur.

Kolnik, Jan↗

Spontaneous sodium ion storage behaviors of reduced graphene oxide anodes exceeding 100% Coulombic efficiency by modulated ion solvation

Rechargeable batteries are essential energy storage devices that power portable devices and electrical vehicles throughout the world. In general, it is thought that the electrochemical performance of rechargeable batteries is mostly determined by the electrodes within them and that the electrolyte plays a relatively passive role. However, ion transport and storage can be greatly influenced by the electrolyte solution structure, specifically, ion solvation within the bulk and ion desolvation across the electrode/electrolyte interfaces. Herein, we studied the role of the electrolyte as an active component of electrochemical energy storage devices. We found that with an appropriate electrolyte formulation, ion storage in disordered carbonaceous anode materials can occur spontaneously without externally supplied electrical energy. Reduced graphene oxide (RGO) in an ether-based electrolyte demonstrates ‘spontaneous' ion storage behaviors of adsorbing and inserting the solvated ions utilizing facilitated permeability and wettability of RGO, which results in Coulombic efficiency of ~145% due to additional charging capacity of ~180 mAh g -1 during electrochemical processes. The unexpected spontaneous ion storage behavior was extensively investigated using a combination of electrochemical analyses and diagnostics, advanced characterizations, and computational simulation. In conclusion, we believe the spontaneous ion storage behavior offers a new way to further improve the energy efficiency of practical rechargeable batteries.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Leveraging dendritic complexity for neuromorphic computing

Abstract Beyond-von Neumann computing approaches are necessary to sustain the growth of microelectronics and the increasing appetite for artificial intelligence/machine learning algorithms. Neuromorphic computing is an emerging paradigm that takes inspiration from the brain to provide a path forward to improve the computational efficiency and computational density of next-generation computing architectures. In nature, we observe brains performing complex computations with a much smaller energy footprint than conventional computing approaches. Current neuromorphic systems are focused primarily on scalability, namely, increasing the number of computational units (neurons) and connections between units (synapses). However, for brain-like cognition and efficiency in next-generation computing hardware, we need increased complexity in function, as well as improved connection density for scalability. Here, we present our work that aims to incorporate dendrites for ‘compute-on-wire’ in neuromorphic architectures to increase the computational complexity (e.g. number of programmable parameters, nonlinear dynamics) as well as computational efficiency (energy/compute) of artificial neural networks (ANNs). We do this by showcasing neuromorphic dendrite elements that can be leveraged for various applications. We will present examples of neuroscience-inspired direction-selective circuits and an ANN with active dendrites leveraging shunting inhibition. We also demonstrate the benefits of using dendrites in deep neural networks. To conclude, we discuss how we can utilize emerging hardware devices in these systems and design next-generation neuromorphic architectures with dendrites.

Cardwell, Suma G. (ORCID:0000000226575545)↗

Addressing Rising Energy Demand Through Innovation

The U.S. is facing a significant increase in energy demand, driven by AI advancements, the rapid expansion of data centers, manufacturing and industrial growth, and the electrification of transportation and buildings. Buildings alone account for approximately 75% of U.S. electricity consumption and 40% of total energy use. To address these challenges, NLR leverages its state-of-the-art research facilities, advanced energy modeling, hardware-in-the-loop emulation, and real-world demonstrations to provide data-driven insights that de-risk emerging energy solutions, increase efficiency and demand flexibility, optimize grid controls, and identify vulnerabilities to enhance energy security. This presentation will highlight our research ecosystem and its role in supporting a more reliable, affordable, and adaptive energy infrastructure in the face of accelerating demand.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Real-Space Constrained Density Functional Theory Investigation of Site-Specific, Interfacial Charge Recombination Dynamics Across the Au Nanoparticle/TiO 2 Heterojunction

Au nanoparticle (NP)/TiO 2 heterojunction is a representative system to study interfacial charge transfer in photocatalysis and photovoltaics, where suppressing recombination from TiO 2 to Au can enhance hot carrier extraction. We apply real-space constrained density functional theory (CDFT) with Marcus theory to quantify charge recombination time scales across Au/TiO 2 . This approach enables direct control and visualization of charge-separated states, aligning with site-specific probes like time-resolved X-ray photoelectron spectroscopy (trXPS). We find that the charge-separated state features a bipolaron, with recombination dominated by TiO 2 LUMO to Au HOMO transitions, primarily at interfacial Au sites. Marcus rate predictions are benchmarked with surface hopping methods, quantifying differences in time scales and computational efficiency. Lastly, we examine how the Au cluster size affects the free energy change (ΔG) and reorganization energy (λ), explaining trends in closed-shell systems and highlighting challenges for open-shell extrapolations. Overall, CDFT + Marcus theory provides efficient, mechanistically transparent interfacial charge transfer modeling, and we clearly defined its applicability and limitation.

Glenna, Drew M. [Univ. of Idaho, Idaho Falls, ID (↗

Economic Analysis of Battery Energy Storage Systems Incorporating Uncertain Battery Model

A high-fidelity battery model is essential for precise economic analysis of battery energy storage systems (BESSs), but these models are computationally intensive. Heuristic models offer computational efficiency but compromise the accuracy of economic analysis results. We assess the impact of errors in heuristic battery models on economic analysis by utilizing open-circuit voltage (OCV) measurements from battery experiments.

Choi, Hyungjin [Sandia National Laboratories (SNL-↗

Gain-enhanced free-electron laser with an electromagnetic pump field

The feasibility of enhancing the gain for the free electron laser (FEL) with an electromagnetic (EM) pump field has been considered. The enhancement is provided by reacceleration of the electrons in the interaction region by means of a static, axial electric field. A FEL utilizing a low energy electron beam and a wiggler magnet with a periodicity on the order of 1 cm would produce far infared (FIR) radiation with wavelengths on the order of a few hundred microns. The use of the FIR radiation as the pump field in a two-stage FEL is envisioned to obtain visible radiation with a low energy electron beam. A summary is provided regarding the theory and equations of motion for the EM-pumped FEL. The derived relations are applied to the second stage of such a two-stage FEL. The obtained equations have been incorporated into a computer code which has been used to calculate laser gain and energy conversion efficiency.

Hiddleston, H. R.↗

Experiments and Model Development for the Investigation of Sooting and Radiation Effects in Microgravity Droplet Combustion

Today, despite efforts to develop and utilize natural gas and renewable energy sources, nearly 97% of the energy used for transportation is derived from combustion of liquid fuels, principally derived from petroleum. While society continues to rely on liquid petroleum-based fuels as a major energy source in spite of their finite supply, it is of paramount importance to maximize the efficiency and minimize the environmental impact of the devices that burn these fuels. The development of improved energy conversion systems, having higher efficiencies and lower emissions, is central to meeting both local and regional air quality standards. This development requires improvements in computational design tools for applied energy conversion systems, which in turn requires more robust sub-model components for combustion chemistry, transport, energy transport (including radiation), and pollutant emissions (soot formation and burnout). The study of isolated droplet burning as a unidimensional, time dependent model diffusion flame system facilitates extensions of these mechanisms to include fuel molecular sizes and pollutants typical of conventional and alternative liquid fuels used in the transportation sector. Because of the simplified geometry, sub-model components from the most detailed to those reduced to sizes compatible for use in multi-dimensional, time dependent applied models can be developed, compared and validated against experimental diffusion flame processes, and tested against one another. Based on observations in microgravity experiments on droplet combustion, it appears that the formation and lingering presence of soot within the fuel-rich region of isolated droplets can modify the burning rate, flame structure and extinction, soot aerosol properties, and the effective thermophysical properties. These observations led to the belief that perhaps one of the most important outstanding contributions of microgravity droplet combustion is the observation that in the absence of asymmetrical forced and natural convection, a soot shell is formed between the droplet surface and the flame, exerting an influence on the droplet combustion response far greater than previously recognized. The effects of soot on droplet burning parameters, including burning rate, soot shell dynamics, flame structure, and extinction phenomena provide significant testing parameters for studying the structure and coupling of soot models with other sub-model components.

Choi, Mun Young↗

A Decision-Support Model for Selecting Additive Manufacturing Versus Subtractive Manufacturing Based on Energy Consumption

This paper presents a simple computational model for determining whether additive manufacturing or subtractive manufacturing is more energy efficient for production of a given metallic part. The key discriminating variable is the fraction of the bounding envelope that contains material – i.e. the volume fraction of solid material. For both the additive process and the subtractive process, the total energy associated with the production of a part is defined in terms of the volume fraction of that part. The critical volume fraction is that for which the energy consumed by subtractive manufacturing equals the energy consumed by additive manufacturing. For volume fractions less than the critical value, additive manufacturing is more energy efficient. For volume fractions greater than the critical value, subtractive manufacturing is more efficient. The model considers the entire manufacturing lifecycle – from production and transport of feedstock material through processing to return of post-production scrap for recycling. Energy consumed by processing equipment while idle is also accounted for in the model. Although the individual energy components in the model are identified and accounted for in the expressions for additive and subtractive manufacturing, values for many of these components may not be currently available. Energy values for some materials’ production and subtractive and additive manufacturing processes can be found in the literature. However, since many of these data are reported for a very specific application, it may be difficult, if not impossible, to reliably apply these data to new process-material manufacturing scenarios since, very often, insufficient information is provided to enable extrapolation to broader use. Consequently, this paper also highlights the need to develop improved knowledge of the energy embodied in each phase of the manufacturing process. To be most valuable, users of the model should determine the energy consumed by their manufacturing process equipment on the basis of energy-per-unit-volume of production for each material of interest – considering both alloy composition and form. Energy consumed during machine idle per unit time should also be determined by the user then scaled to specific processing scenarios. Energy required to generate feedstock material (billet, plate, bar, wire, powder) must be obtained from suppliers.

J K Watson↗