Search NASA⌕ Search

SEARCH · Search NASA

Results for “efficient”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 379 records · Page 21

Toward Energy-Efficient HPC: Insights from Power Profiling a Cloud-Resolving Earth System Model

Power is a fundamental constraint as supercomputing advances to exascale. Efficient operation within strict power budgets requires application-aware power management based on a detailed understanding of application-level power behavior. This work analyzes the Energy Exascale Earth System Model (E3SM) atmosphere component, SCREAM, on Perlmutter (NERSC) and Frontier (OLCF). We characterize power variation across inputs, concurrency levels, and power caps, evaluate the energy impact of code optimizations, and attribute energy within the code using a newly developed GPU energy model. Results show that SCREAM’s peak power remains stable during its core execution phase and decreases gradually as concurrency increases. Power capping experiments reveal a performance–energy "sweet spot". On Perlmutter, limiting GPU power to 50% of thermal design power (TDP) achieves up to 15% energy savings with a 7% performance penalty. On Frontier, a 40% TDP cap yields up to 10% energy savings with less than 10% performance loss. Code optimizations reduce SCREAM energy by shortening run time without increasing power. Modeling reveals a critical insight: data movement accounts for approximately 70% of SCREAM’s GPU energy. This fundamentally shifts the optimization focus from FLOPS to data transfer reduction for this class of applications, offering the most impactful strategy for improving energy efficiency. This work establishes a foundation for practical, application-aware power management at exascale.

Zhao, Zhengji [Lawrence Berkeley National Laborato↗

15.3% AM1.5G Efficiency GaAs Solar Cells Fabricated via an Epitaxy-Free Process

Here, we report simple and potentially low-cost techniques for creating high-quality n-type gallium arsenide (GaAs) and GaAs p/n junctions and fabricate GaAs p/n junction solar cells. Detailed-balance modeling suggests that 20% AM1.5G efficiency p/n homojunction devices may be possible if the surface doping concentration can be limited to values less than ∼ 5 × 10 19 cm −3 . Our process exploits an open-tube, vapor-phase, deposition-free, zinc diffusion technique for forming p-type layers in melt-grown n-GaAs substrates that results in sheet resistances less than 1 kΩ/$\square$. In addition, we have improved the minority carrier diffusion lengths of melt-grown GaAs from less than one micron to over five microns using an open-tube, vacuum-free, annealing process which reduces the density of EL2 midgap defects. Finally, we have combined these advances to fabricate epitaxy-free, GaAs solar cells with a validated AM1.5G efficiency of 15.3%.

14 SOLAR ENERGY↗

Slope Efficiency and Voltage Reduction at High Current Densities in AlInGaAs Diode Lasers

Here, the slope efficiency and drive voltage of broad area AlInGaAs laser diodes near 865 nm is observed to decrease significantly under quasi-CW pulsed operation at currents well above threshold, in a manner that cannot be explained by thermal effects or carrier leakage over heterojunction barriers. Simulations show that the slope efficiency reduction is explicable by increased free carrier absorption in the waveguide region. Empirical formulas are presented to represent these effects in a closed analytic form suitable for use in simulators for diode-pumped laser systems.

47 OTHER INSTRUMENTATION↗

Efficient Streaming Dynamic Mode Decomposition

We propose a reformulation of the streaming dynamic mode decomposition method that requires maintaining a single orthonormal basis, thereby reducing computational redundancy. The proposed efficient streaming dynamic mode decomposition method results in a constant-factor reduction in computational complexity and memory storage requirements. Numerical experiments on representative canonical dynamical systems show that the enhanced computational efficiency does not compromise the accuracy of the proposed method.

97 MATHEMATICS AND COMPUTING↗

Early Exploration of a Flexible Framework for Efficient Quantum Linear Solvers in Power Systems

The rapid integration of renewable energy resources presents formidable challenges in managing power grids. While advanced computing and machine learning techniques offer some solutions for accelerating grid modeling and simulation, there remain complex problems that classical computers cannot effectively address. Quantum computing, a promising technology, has the potential to fundamentally transform how we manage power systems, especially in scenarios with a higher proportion of renewable energy sources. One critical aspect is solving linear systems of equations, crucial for power system applications like power flow analysis, for which the Harrow-Hassidim-Lloyd (HHL) algorithm is a well-known quantum solution. However, HHL quantum circuits often exhibit excessive depth, making them impractical for current Noisy-Intermediate-Scale-Quantum (NISQ) devices. In this paper, we introduce a versatile framework, powered by NWQSim, that bridges the gap between power system applications and quantum linear solvers available in Qiskit. This framework empowers researchers to efficiently explore power system applications using quantum linear solvers. Through innovative gate fusion strategies, reduced circuit depth, and GPU acceleration, our simulator significantly enhances resource efficiency. Power flow case studies have demonstrated up to a eight-fold speedup compared to Qiskit Aer, all while maintaining comparable levels of accuracy.

quantum computing, Harrow-Hassidim-Lloyd, high-per↗

Optimizing Grid-interactive Efficient Building Designs with Stacked Value Streams

Grid-interactive efficient buildings (GEBs) are those characterized by the combination of energy efficiency and demand flexibility with smart technologies and communications to not only deliver greater affordability and comfort to buildings, but also help utilities manage grid operations and lower system costs. This paper presents an innovative techno-economic assessment framework to effectively examine different GEB design options, explore various use cases, define technically achievable benefits, and thereby assist in informed decision-making. In particular, building load flexibility, thermal storage, and battery energy storage are considered. Advanced optimal dispatch problem is formulated to maximize the stacked value streams from multiple, competing use cases, subject to the physical capabilities and operational flexibility associated with different designs and configurations. Comprehensive case studies were performed for a real-world building to evaluate the cost-effectiveness of different GEB designs and offer in-depth insights. It was found that the proposed assessment method could effectively capture the costs and benefits linked to each GEB design option. Furthermore, the study revealed that outage mitigation and demand response are the two most significant sources of benefits for GEBs.

Ma, Xu↗

Model-Free Control of Grid-Interactive Efficient Buildings Under Communication Time Delays

Grid-interactive efficient buildings (GEBs) have recently been used to enhance the reliability and stability of the electric grid through demand response (DR) programs. However, most existing DR control strategies require accurate modeling of the various building thermostatically controlled loads (TCLs) and are computationally expensive. To address these challenges, a model-free control (MFC)-based strategy has recently been introduced for coordinating and controlling GEBs. MFC is a data-enabled control strategy that is computationally efficient and does not require the analytical models of the various building equipment. In this paper, we numerically investigate the impact of communication time delays on the performance of MFC in maintaining the TCLs' temperatures within the desired comfort levels while meeting the assigned power allocation constraint.

Telsang, Bhagyashri [University of Tennessee, Knox↗

Single-Cell Universal Logic-in-Memory Using 2T-nC FeRAM: An Area and Energy-Efficient Approach for Bulk Bitwise Computation

This work presents a novel approach to configure 2T-nC ferroelectric RAM (FeRAM) for performing single cell logic-in-memory operations, highlighting its advantages in energy-efficient computation over conventional DRAM-based approaches. Unlike conventional 1T-1C dynamic RAM (DRAM), which incurs refresh overhead, 2T-nC FeRAM offers a promising alternative as a non-volatile memory solution with low energy consumption. Our key findings include the potential of quasi-nondestructive readout (QNRO) sensing in 2T-nC FeRAM for logic-in-memory (LiM) applications, demonstrating its inherent capability to perform inverting logic without requiring external modifications, a feature absent in traditional 1T-1C DRAM. We successfully implement the MINORITY function within a single cell of 2T-nC FeRAM, enabling universal NAND and NOR logic, validated through SPICE simulations and experimental data. Additionally, the research investigates the feasibility of 3D integration with 2T-nC FeRAM, showing substantial improvements in storage and computational density, facilitating bulk-bitwise computation. Our evaluation of eight real-world, data-intensive applications reveals that 2T-nC FeRAM achieves 2× higher performance and 2.5× lower energy consumption compared to DRAM. Furthermore, the thermal stability of stacked 2T-nC FeRAM is validated, confirming its reliable operation when integrated on a compute die. These findings emphasize the advantages of 2T-nC FeRAM for LiM, offering superior performance and energy efficiency over conventional DRAM.

36 MATERIALS SCIENCE↗

Efficient Simulation of Cascading Outages Using an Energy Function-Embedded Quasi-Steady-State Model

Here, this paper proposed an energy function-embedded quasi-steady-state model for efficient simulation of cascading outages on a power grid while addressing transient stability concerns. Compared to quasi-steady-state models, the proposed model incorporates short-term dynamic simulation and an energy function method to efficiently evaluate the transient stability of a power grid together with outage propagation without transient stability simulation. Cascading outage simulation using the proposed model conducts three steps for each disturbance such as a line outage. First, it performs time-domain simulation for a short term to obtain a post-disturbance trajectory. Second, along the trajectory, the system state with the local maximum potential energy is found and used as the initial point to search for a relevant unstable equilibrium by Newton's method. Third, the transient energy margin is estimated based on this unstable equilibrium to predict an out-of-step condition with generators. The proposed energy function-embedded quasi-steady-state model is tested in terms of its accuracy and time performance on an NPCC 140-bus power system and compared to a quasi-steady-state model embedding transient stability simulation.

Guo, Zhenping [Univ. of Tennessee, Knoxville, TN (↗

Size-related decline in dryland shrubs is related to reductions in hydraulic efficiency and carbon assimilation and not nonstructural carbohydrate depletion

Plant growth and survival are fundamentally constrained by water transport from roots to leaves, impacting carbon assimilation and associated labile carbon pools. However, physiological constraints on growth and survival vary with plant age, due to changes in metabolic sinks, and increases in hydraulic path length from rhizosphere to canopy. We investigated crown dieback, growth, hydraulics, carbon assimilation and carbohydrate storage in relation to increasing basal diameter of two dominant shrub species (Caragana korshinskii and Artemisia ordosica) at the southeastern edge of the Tengger Desert, China. The aim was to identify mechanisms of decreased performance with plant size in dryland shrubs. Clear contrasts in stomatal regulation of leaf water potentials were detected between species. Despite these contrasts, radial growth, hydraulic transport efficiency (Ks), and photosynthetic efficiency similarly declined in both species with increasing plant size, while non-structural carbohydrate (NSC) reserves remained unchanged. Xylem embolism (PLC) increased with plant size, resulting in significant reductions in carbon assimilation in both species. Results indicate that hydraulic, and potentially carbon assimilation constraints, rather than reductions in carbohydrate storage, govern growth-related dryland shrub decline. These findings improve our understanding of how population demography impacts dryland forest response to climate change.

Zhang, Hongxia↗

Leveraging data mining, active learning, and domain adaptation for efficient discovery of advanced oxygen evolution electrocatalysts

Developing advanced catalysts for acidic oxygen evolution reaction (OER) is crucial for sustainable hydrogen production. This study presents a multistage machine learning (ML) approach to streamline the discovery and optimization of complex multimetallic catalysts. Our method integrates data mining, active learning, and domain adaptation throughout the materials discovery process. Unlike traditional trial-and-error methods, this approach systematically narrows the exploration space using domain knowledge with minimized reliance on subjective intuition. Then, the active learning module efficiently refines element composition and synthesis conditions through iterative experimental feedback. The process culminated in the discovery of a promising Ru-Mn-Ca-Pr oxide catalyst. Our workflow also enhances theoretical simulations with domain adaptation strategy, providing deeper mechanistic insights aligned with experimental findings. By leveraging diverse data sources and multiple ML strategies, we demonstrate an efficient pathway for electrocatalyst discovery and optimization. This comprehensive, data-driven approach represents a paradigm shift and potentially benchmark in electrocatalysts research.

Science & Technology - Other Topics↗

Two-dimensional perovskite templates for durable, efficient formamidinium perovskite solar cells

We present a design strategy for fabricating ultrastable phase-pure films of formamidinium lead iodide (FAPbI 3 ) by lattice templating using specific two-dimensional (2D) perovskites with FA as the cage cation. When a pure FAPbI 3 precursor solution is brought in contact with the 2D perovskite, the black phase forms preferentially at 100°C, much lower than the standard FAPbI 3 annealing temperature of 150°C. X-ray diffraction and optical spectroscopy suggest that the resulting FAPbI 3 film compresses slightly to acquire the (011) interplanar distances of the 2D perovskite seed. The 2D-templated bulk FAPbI 3 films exhibited an efficiency of 24.1% in a p-i-n architecture with 0.5–square centimeter active area and an exceptional durability, retaining 97% of their initial efficiency after 1000 hours under 85°C and maximum power point tracking.

14 SOLAR ENERGY↗

An Optimization-Based Coupling of Reduced Order Models with an Efficient Reduced Adjoint Basis Generation Approach

Optimization-based coupling (OBC) is an attractive alternative to traditional Lagrange multiplier approaches in multiple modeling and simulation contexts. However, application of OBC to time-dependent problems has been hindered by the computational cost of finding the stationary points of the associated Lagrangian, which requires primal and adjoint solves. This issue can be mitigated by using OBC in conjunction with computationally efficient reduced order models (ROMs). To demonstrate the potential of this combination, in this paper, we develop an optimization-based ROM-ROM coupling for a transient advection-diffusion transmission problem. We pursue the “optimize-then-reduce” path toward solving the minimization problem at each time step and solve reduced space adjoint system of equations, where the main challenge in this formulation is the generation of adjoint snapshots and reduced bases for the adjoint systems required by the optimizer. One of the main contributions of the paper is a new technique for an efficient adjoint snapshot collection for gradient-based optimizers in the context of optimization-based ROM-ROM couplings. In conclusion, we present numerical studies demonstrating the accuracy of the approach along with comparison between various approaches for selecting a reduced order basis for the adjoint systems, including decay of snapshot energy, average iteration counts, and timings.

coupled problems↗

Efficient CP Rounding Using Alternating Least Squares with QR Decomposition

The CANDECOMP/PARAFAC (CP) decomposition is widely used for analyzing multidimensional data, and the alternating least squares (CP-ALS) algorithm is a common method for its computation. CP rounding is the problem of computing a lower-rank CP decomposition of an input already in a higher-rank CP format. While the normal equations (NE) approach in CP-ALS is efficient for the CP rounding problem and frequently used, it becomes unstable in the presence of ill-conditioned subproblems. This paper presents a new QR-based CP-ALS method for CP rounding that preserves both numerical stability and computational efficiency. Here, our experiments show that the proposed method offers significant speedup over a previous QR-based approach and the Tensor Toolbox's NE-based implementation, particularly for higher-order tensors. Furthermore, our approach demonstrates a marked reduction in error for ill-conditioned problems, with error reductions several orders of magnitude smaller compared to the NE-based method, while achieving faster convergence and more accurate solutions. By using a more numerically stable approach, we can solve more problems in reduced working precision, which enables further reduction in time to solution.

CANDECOMP/PARAFAC↗

AI-Powered Knowledge Graphs for Neuromorphic and Energy-Efficient Computing

The surge in scientific literature obscures breakthroughs and hinders the discovery of new research paths. We propose an artificial intelligence (AI) powered framework using large language models (LLMs) and knowledge graphs (KGs) to automate parts of scientific discovery, focusing on energy-efficient AI circuits. Our hybrid approach combines LLMs, structured data, and ontology-based reasoning to construct a comprehensive knowledge graph that integrates insights across computational neuroscience, spiking neuron models, learning rules, architectural motifs, and neuromorphic device technologies. This multi-domain representation enables the generation of hypotheses that connect biological function with implementable, energy-efficient hardware architectures. Using KG embeddings and graph neural networks, the framework generates hypotheses for novel circuits, validates them through optimization on exascale HPC systems, and with tools like SuperNeuro and Fugu, the most promising designs will be prototyped in hardware. This open-source system aims to accelerate discoveries and bridging neuroscience with hardware innovation, drive collaboration, and unlock new opportunities in low-power AI computing.

Gautam, Ashish [ORNL]↗

Intelligent Sampling of Extreme-Scale Turbulence Datasets for Accurate and Efficient Spatiotemporal Model Training

With the end of Moore’s law and Dennard scaling, efficient training increasingly requires rethinking data volume. Can we train better models with significantly less data via intelligent subsampling? To explore this, we develop SICKLE, a sparse intelligent curation framework for efficient learning, featuring a novel maximum entropy (MaxEnt) sampling approach, scalable training, and energy benchmarking. We compare MaxEnt with random and phase-space sampling on large direct numerical simulation (DNS) datasets of turbulence. Evaluating SICKLE at scale on Frontier, we show that subsampling as a preprocessing step can, in many cases, improve model accuracy and substantially lower energy consumption, with observed reductions of up to 38×.

Brewer, Wes [ORNL] (ORCID:0000000236393956)↗

JANUS: Resilient and Adaptive Data Transmission for Enabling Timely and Efficient Cross-Facility Scientific Workflows

In modern science, the growing complexity of large-scale scientific projects has led to an increasing reliance on cross-facility scientific workflows, where resources and expertise from multiple institutions and geographic locations are leveraged to accelerate scientific discovery. These workflows often require transmitting huge amounts of scientific data through wide-area networks. Although high-speed networks like ESnet and transfer services such as Globus have improved data mobility, several challenges remain. The sheer volume of data can overwhelm network bandwidth, widely used transport protocols such as TCP suffer from inefficiencies due to retransmissions triggered by packet loss, and existing fault-tolerance mechanisms like erasure coding introduce substantial overhead. In this paper, we propose Janus, a resilient and adaptable data transmission approach designed for cross-facility scientific workflows. Unlike traditional TCP-based methods, Janus leverages UDP, integrates erasure coding for fault tolerance, and combines it with error-bounded lossy compression to reduce overhead. This novel design allows users to balance data transmission time and accuracy, optimizing transfer performance based on specific scientific requirements. Additionally, Janus dynamically adjusts erasure coding parameters in response to real-time network conditions, ensuring efficient data transfers even in fluctuating environments. We develop optimization models for determining ideal configurations and implement adaptive data transfer protocols to enhance reliability. Through extensive simulations and real-network experiments, we demonstrate that Janus significantly improves transfer efficiency while maintaining data fidelity.

Esaulov, Vladislav [Georgia State University, Atla↗

Identifying Critical Electrode Metrics for Efficient, Selective CO 2 Electrochemical Conversion

Low-temperature electrochemical CO 2 reduction (CO 2 R) in zero-gap membrane electrode assembly (MEA) reactors presents a scalable route to fuels and carbon utilization. However, performance at industrially relevant current densities hinges on mesoscale catalyst layer integration, particularly at the ionomer|catalyst interface. Here, we demonstrate a generalizable in situ electrochemical impedance spectroscopy (EIS) method. We utilize this technique to decouple electrode-level parameters that are correlated to the overall MEA performance. By performing this ex situ EIS method on CO 2 -to-CO catalyst-coated membranes with systematically varied ionomer-to-catalyst (I:C) ratios, we reveal a pronounced dependence of performance, ion transport resistance, and catalyst utilization on the I:C ratio as well as the electrode conditioning. We demonstrate that an optimal I:C ratio exists at which ion transport resistance is minimized and Faradaic efficiency for CO production is maximized. Beyond the electrodes examined, here we compare ion transport resistance to MEA selectivity/Faradaic efficiency obtained in prior studies, revealing a clear correlation between the two. These results suggest that ion transport resistance within the catalyst layer may be a quantitative predictor of MEA performance which underscores the importance of mesoscale integration in achieving scalable CO 2 R technologies.

08 HYDROGEN↗