Search NASA⌕ Search

SEARCH · Search NASA

Results for “energy efficient computing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 235 records · Page 13

Next-generation tunnel FETs: exploring material perspectives and areal tunneling configurations

The end of Dennard scaling, which facilitated proportional increases in computing power without added energy costs until the mid-2000s, has underscored the urgent need for innovative semiconductor devices that can enhance energy efficiency. Tunnel field-effect transistors (TFETs) have emerged as promising candidates to surpass the energy efficiency of conventional metal oxide semiconductor field-effect transistors (MOSFETs). Unlike MOSFETs, which rely on thermionic emission to overcome the source-channel potential barrier, TFETs operate through quantum tunneling, potentially enabling sub-60 mV dec −1 subthreshold swing (SS) for low-voltage operation. However, lateral TFETs have faced challenges in achieving adequate on-state current (I ON ) and a broad SS operation window, limiting their practical utility. This review article advocates for areal TFETs, which utilize face-to-face tunnel junctions that ideally offer step-function current turn-on characteristics and allow I ON to scale with device area rather than width. We highlight recent advancements in integrating 2D materials into tunneling structures, which could facilitate efficient band-to-band tunneling through atomically thin layers, while addressing challenges of gate field screening. We then discuss the nearer-term prospects of epitaxial areal TFETs comprising III–V compound semiconductors and group-IV semiconductors based on recent experimental progress. The review examines both quantum mechanical and semiclassical modeling approaches for TFETs, including techniques to reduce the computational complexity. The article delves into ongoing challenges in material synthesis, interface engineering, device fabrication, and integration pathways, concluding with recommendations for future research directions to overcome the fundamental power density limitations of conventional transistor technology.

2D materials↗

Combining Generative Modeling and Advanced Control for Building Scenario Generation

Buildings make up a large portion of energy consumption in the U.S. today. Understanding their energy consumption patterns can improve their efficiency, but requires detailed models that rely on incomplete or unknown information. Previous work has shown that artificial intelligence (AI) can be used to predict missing information and even suggest upgrades to improve building efficiency. However, building upgrades may require undesirable upfront costs. Oppositely, advanced control could improve building efficiency with negligible upfront cost. To explore the tradeoffs between these two approaches, in this work we propose a workflow to compute optimal temperature setpoint schedules to minimize energy consumption and operational cost. Results show that modifying the temperature setpoints in a building using model predictive control (MPC) can effectively reduce its energy consumption and operational cost. This optimal operation cannot fully meet a desired goal. However, we show that by considering MPC in addition to component upgrades, a desired goal can be met with significantly less upfront costs.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Mantaray: A Rust Package for Ray Tracing Ocean Surface Gravity Waves

Ocean surface gravity waves are an important component of air-sea interaction, influencing energy, momentum, and gas exchanges across the ocean-atmosphere interface. In specific applications such as refraction by ocean currents or bathymetry, ray tracing provides a computationally efficient way to gain insight into wave propagation. In this paper, we introduce Mantaray, an open-source software package implemented in Rust, with a Python interface, that solves the ray equations for ocean surface gravity waves. Mantaray is designed for performance, robustness, and ease of use. The package is modular to facilitate further development and can currently be applied to both idealized and realistic wave propagation problems (Fig. 1).

16 TIDAL AND WAVE POWER↗

Dynamical structure factors of warm dense matter from time-dependent orbital-free and mixed-stochastic-deterministic density functional theory

Abstract We present the first calculations of the inelastic part of the dynamical structure factor (DSF) for warm dense matter (WDM) using time-dependent orbital-free density functional theory (TD-OF-DFT) and mixed-stochastic-deterministic (mixed) Kohn Sham TD-DFT (KS TD-DFT). WDM is an intermediate phase of matter found in planetary cores and laser-driven experiments, where the accurate calculation of the DSF is critical for interpreting x-ray Thomson scattering measurements. Traditional TD-DFT methods, while highly accurate, are computationally expensive, motivating the exploration of TD-OF-DFT and mixed TD-KS-DFT as more efficient alternatives. We applied these methods to experimentally measured WDM systems, including solid-density aluminum and beryllium, compressed beryllium, and carbon–hydrogen mixtures. Our results show that TD-OF-DFT requires a dynamical kinetic energy potential in order to qualitatively capture the plasmon response. Additionally, it struggles with capturing bound electron contributions. In contrast, mixed TD-KS-DFT offers greater accuracy in distinguishing bound and free electron effects, aligning well with experimental data, though at a higher computational cost. This study highlights the trade-offs between computational efficiency and accuracy, demonstrating that TD-OF-DFT remains a valuable tool for rapid scans of parameter space, while mixed TD-KS-DFT should be preferred for high-fidelity simulations. Our findings provide insight into the future development of DFT methods for WDM and suggest potential improvements for TD-OF-DFT.

36 MATERIALS SCIENCE↗

Universal reduced basis for the calibration of covariant energy density functionals

The reduced basis method is used to construct a “universal” basis of Dirac orbitals that may be applicable throughout the nuclear chart to calibrate covariant energy density functionals. Relative to the successful development of a reduced basis emulator for the nonrelativistic Schrödinger equation, the Dirac equation adds an extra layer of complexity due to the existence of negative energy states, which complicates building an efficient reduced basis. However, once this problem is mitigated, the resulting reduced basis is able to accurately and efficiently reproduce the high-fidelity model at a fraction of the computational cost. We are confident that the resulting reduced basis will serve as a foundational element in developing rapid and accurate emulators. In turn, these emulators will play a critical role in the Bayesian optimization of covariant energy density functionals.

Bayesian methods↗

HPE ultralit project (ARPA-E open program 2018) final report 15

The goal of the program was to build a fully integrated optical transceiver with >1 Tb/s and <1.5 pJ/bit operating at 50 ◦ C. Optical transceivers are critical components in high-performance-computers (HPC) and data centers, and the details of their implementation has a big impact on the total power consumption (energy efficiency) of an HPC system. Our proposed transceiver used three key enabling technologies. Firstly, SiGe avalanche photodetectors have record-low sensitivities, meaning that they can reach low bit error rates with very little light input. As a result, we can drive our lasers at a lower drive current, thus saving electrical power. Secondly, we use MOS-based capacitive tuning in our deinterleaver, modulator, and demultiplexer. Capacitive tuning allows for the tuning of photonic elements with zero static power consumption. Thirdly, we use quantum dots as the gain material in our light source. This allows us to efficiently use our light source at temperatures that are typically encountered in an HPCsystem. Our proposed optical transceiver consisted of a quantum dot comb laser as a light source, a booster SOA, MOS-based deinterleavers, MOS based ring modulators, MOS-based ring demultiplexers, and SiGe APDs.

77 NANOSCIENCE AND NANOTECHNOLOGY↗

2024 Second Half Semi Annual Report: Modeling plasticity-mediated flow in metals with pressurized cavities

The objective is to better predict the bulk-scale mechanical behavior of porous metals that have over pressurized cavities (e.g., irradiated metals with helium bubbles) by quantifying the complex coupling among cavity aspects (e.g., size distribution, inhomogeneous overpressure values, spatial arrangement) and metal properties (e.g., rate-dependency, crystallographic lattice). This requires up-scaling local mechanical fields from the single crystal scale and will be accomplished using a homogenization approach that combines full-field numerical simulations, analytical formalisms, and physics-informed machine learning to produce symbolically-defined constitutive equations (e.g., gauge functions). These equations will satisfy the objective because they enable computationally efficient predictions that approach the accuracy of computationally expensive full-field numerical simulations, abide by theoretical requirements (e.g., conservation of energy, work conjugacy), and retain the transparency of analytical models.

36 MATERIALS SCIENCE↗

Alfalfa Virtual Building Service: Software Engineering Best Practices Applied to Runtime Interaction with Building Energy Models

Buildings are active participants in increasingly complex energy systems. Building Energy Modeling (BEM) has a key role to play in planning and de-risking an equitable energy transition, with BEM-backed "virtual buildings" critical path for diverse applications that include workforce training tools, Hardware-in-the-Loop (HIL) experimentation to study equipment performance under a range of conditions, Control-Hardware-in-the-Loop (CHIL) experimentation to de-risk commercial control implementations at equipment through grid orchestration levels, and integration of dynamic load profiles into grid modeling tools for energy system experimentation at the urban scale. Modeling requirements vary across these applications, but many software engineering tasks do not. The Alfalfa Virtual Building Service (AVBS, see https://github.com/NREL/alfalfa/wiki) is an open-source web service that solves these common tasks robustly in one place, providing a foundational platform for power users to bootstrap their own applications. AVBS abstracts the specifics of runtime interaction with OpenStudio, Modelica, and Spawn of EnergyPlus models behind a unified REST API. Additionally, AVBS provides resources for cloud deployment and scaling to 100s of parallel simulations, a growing library of modular Operational Technology (OT) integrations for emulation of real-world interfaces, and scripts to automate the population of communities of virtual buildings from URBANopt, ResStock and ComStock.

building automation↗

ReEDS Performance Improvement

The Regional Energy Deployment System (ReEDS) is an open-source, spatially explicit, long-term capacity expansion model for the bulk electric power system of the contiguous United States, encompassing multiple scenarios with technological and political assumptions (see https://github.com/NREL/ReEDS-2.0). With the increased needs for capabilities, higher temporal and spatial resolutions to model the evolution of the power system with modern technologies and low-carbon pathways, ReEDS' model solution times have increased significantly from 4-6 hours in 2018 to 18-48+ hours in 2023 . Also, the model size for commonly-run ReEDS scenarios reached 22 and 28 million equations and variables, respectively. These runtimes can be especially challenging under certain scenario settings (e.g., very high temporal or spatial resolution) or with limited computational power. In this presentation, we will discuss several methods we used to improve model runtime, including data preparation, model modification, and solver tuning. The implementation of these methods shrank the model size to 7.2 and 7.3 million equations and variables, respectively. Furthermore, this led to a 77% reduction in the model's run time for commonly-run ReEDS scenarios. We will discuss the process of identifying areas for solve time improvements and how the specific enhancements for the ReEDS model might be applied to other similar large-scale models.

ENERGY PLANNING, POLICY, AND ECONOMY,MATHEMATICS A↗

Fundamental microscopic properties as predictors of large-scale quantities of interest: Validation through grain boundary energy trends

Correlations between fundamental microscopic properties computable from first principles, which we term canonical properties, and complex large-scale quantities of interest (QoIs) provide an avenue to predictive materials discovery. Here, we propose that such correlations can be efficiently discovered through simulations utilizing approximate interatomic potentials (IPs), which serve as an ensemble of “synthetic materials”. As a proof of principle we build a regression model relating canonical properties to the symmetric tilt grain boundary (GB) energy curves in face-centered cubic crystals, characterized by the scaling factor in the universal lattice matching model of Runnels et al. (2016), which we take to be our QoI. Our analysis recovers known correlations of GB energy to other properties and discovers new ones. We also demonstrate, using available density functional theory (DFT) GB energy data, that the regression model constructed from IP data is consistent with DFT results, confirming the assumption that the IPs and DFT belong to same statistical pool and thereby validating the approach. Regression models constructed in this fashion can be used to predict large-scale QoIs based on first-principles data and provide a general method for training IPs for QoIs beyond the scope of first-principles calculations.

36 MATERIALS SCIENCE↗

Operation and Control of Electric Vehicle Charger with Enhanced Dynamic Performance Under Non-Ideal Grid Voltage Condition

This paper presents a three-phase electric vehicle charger connected to the grid, featuring multiple boost converters on the DC side, specifically designed to ensure smooth, oscillation-free power transfer during unsymmetrical voltage sags. Precise control mechanisms are implemented on the boost converter side to regulate both voltage and current on the electric vehicle side, thereby maintaining optimal charging conditions. The control architecture for both the grid-connected and boost converter components is based on the Lyapunov energy function, which is employed to achieve superior dynamic performance and stability. The system's robustness and reliability are demonstrated through its ability to maintain stable operation and efficient power transfer despite fluctuations in grid conditions. Furthermore, the implementation of Lyapunov-based control ensures rapid response and minimal energy loss, enhancing the overall efficiency of the system. To validate the effectiveness of this approach, a comprehensive model of the system was developed and tested using MATLAB/Simulink, with detailed computer simulations conducted across various significant case studies.

DC-DC boost converter↗

Comparison of a Full-Scale and a 1:10 Scale Low-Speed Two-Stroke Marine Engine Using Computational Fluid Dynamics

International marine shipping is a growing component of international trade; a vast majority of all the world’s goods are being transported on large ocean-going vessels. The International Maritime Organization (IMO) introduced the Energy Efficiency Design Index in 2013, a regulatory framework of associated metrics for reducing emissions of CO 2 per tonne-mile from shipping by approximately 10% each decade. Therefore, decarbonizing the maritime sector requires the development of new fuel sources. Because of the extremely large physical size of the internal combustion engines present in shipping vessels, experimental iterative development of the engine and fuel system is cost-prohibitive. Thus, the ability to perform combustion system development in a scaled platform that can be more easily operated and modeled computationally is of interest. To that end, scaling relationships are needed to translate the results from a smaller engine to a larger counterpart. Scaling studies to date have been restricted to low scaling ratios, four-stroke light-duty engines, and under-resolved computational fluid dynamic simulations that likely do not accurately capture the physics of scaling. In this work, computational models of a 1:10 scale and a full-scale two-stroke crosshead low-speed marine engine were created and validated against experiments obtained in a real 1:10 scale engine installed at Oak Ridge National Laboratory. Further, due to the large size of the full-scale engine, the model required large high-performance computing resources to be evaluated. The availability of high-performance computing resources at the Department of Energy’s Leadership Computing Facilities is an enabler of the current work. The results of the small- and large-scale engine simulations were compared to analyze the effectiveness of the appropriate scaling laws under these extreme scaling ratio conditions.

33 ADVANCED PROPULSION SYSTEMS↗

Enriching OpenStreetMap network data for transportation applications: Insights into the impact of urban congestion on accessibility

OpenStreetMap (OSM) data is a valuable open-source resource for various transportation, traffic, and planning applications. However, OSM network data lack operating traffic speed information, which is critical for transport planning and operations. Addressing this shortcoming, this study leverages commercial vendor data (to serve as ground truth) with exogenous, open-source variables characterizing local transport infrastructure, land use, and demographic information to predict average congested traffic speeds on OSM networks. Three machine-learning models were tested and estimated for OSM links with and without speed limit information in the Denver metropolitan region. Among these, XGBoost performed best, with mean absolute errors of 3.27 and 3.62 mph for links with and without speed limits, respectively. The developed models accurately predicted traffic speeds for different hours and days of the week compared to ground truth data. Using these predicted speeds, drive accessibility scores were computed for the Denver region for different time periods using the Mobility Energy Productivity (MEP) metric to understand the impact of congestion on energy-efficient accessibility. Results show that congestion-adjusted drive accessibility can be significantly lower compared to accessibility calculated using free flow speeds. Specifically, weekday evening hours saw a 42 % drop in accessibility due to reduced speeds, particularly around downtown Denver. Across the Denver metro region, approximately half as many opportunities and jobs are accessible in under 20 min by car during the evening peak period relative to free flow conditions. These findings underscore the importance of using congestion-adjusted operating speeds rather than speed limits in accessibility calculations, as reliance on speed limits can substantially overestimate energy-efficient drive accessibility in large, car-centric cities susceptible to significant congestion. In conclusion, the methodology presented here could further enrich OSM network data, making them useful for an even broader range of transportation applications.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Unsupervised atomic data mining via multi-kernel graph autoencoders for machine learning force fields

Constructing a chemically diverse dataset while avoiding sampling bias is critical to training efficient and generalizable force fields. However, in computational chemistry and materials science, many common dataset generation techniques are prone to oversampling regions of the potential energy surface. Furthermore, these regions can be difficult to identify and isolate from each other or may not align well with human intuition, making it challenging to systematically remove bias in the dataset. While traditional clustering and pruning (down-sampling) approaches can be useful for this, they can often lead to information loss or a failure to properly identify distinct regions of the potential energy surface due to difficulties associated with the high dimensionality of atomic descriptors. In this work, we introduce the Multi-kernel Edge Attention-based Graph Autoencoder (MEAGraph) model, an unsupervised approach for analyzing atomic datasets. MEAGraph combines multiple linear kernel transformations with attention-based message passing to capture geometric sensitivity and enable effective dataset pruning without relying on labels or extensive training. Demonstrated applications on niobium, tantalum, and iron datasets show that MEAGraph efficiently groups similar atomic environments, allowing for the use of basic pruning techniques for removing sampling bias. This approach provides an effective method for representation learning and clustering that can be used for data analysis, outlier detection, and dataset optimization.

Materials science↗

Scalable training of trustworthy and energy-efficient predictive graph foundation models for atomistic materials modeling: a case study with HydraGNN

We present our work on developing and training scalable, trustworthy, and energy-efficient predictive graph foundation models (GFMs) using HydraGNN, a multi-headed graph convolutional neural network architecture. HydraGNN expands the boundaries of graph neural network (GNN) computations in both training scale and data diversity. It abstracts over message passing algorithms, allowing both reproduction of and comparison across algorithmic innovations that define nearest-neighbor convolution in GNNs. This work discusses a series of optimizations that have allowed scaling up the GFMs training to tens of thousands of GPUs on datasets consisting of hundreds of millions of graphs. Our GFMs use multitask learning (MTL) to simultaneously learn graph-level and node-level properties of atomistic structures, such as energy and atomic forces. Using over 154 million atomistic structures for training, we illustrate the performance of our approach along with the lessons learned on two state-of-the-art US Department of Energy (US-DOE) supercomputers, namely the Perlmutter petascale system at the National Energy Research Scientific Computing Center and the Frontier exascale system at Oak Ridge Leadership Computing Facility. The HydraGNN architecture enables the GFM to achieve near-linear strong scaling performance using more than 2000 GPUs on Perlmutter and 16,000 GPUs on Frontier.

97 MATHEMATICS AND COMPUTING↗

Optimization of Scrap Melting Using an Electric Arc in Steel Manufacturing

Steel industry is crucial to the national economy and security. Around 67% of crude steel in the U.S is produced in electric arc furnaces (EAF), which is energy intensive. Around 140 EAFs operate in the U.S., consuming about 8.6x10 7 MMBtu/year of electricity. One of major challenges for EAFs includes maximizing the efficiency of the electrical energy provided in the form of electric arcs to melt various scrap mixes. To address this issue, a computational fluid dynamics (CFD) methodology is chosen to analyze scrap melting using the electric arc. Due to complex furnace phenomena and the wide variety of potential scenarios, high performance computing (HPC) is essential to yield comprehensive and detailed CFD analyses and systematic parametric studies for optimized EAF operation. The objectives are to 1) simulate scrap melting using electric arc, 2) evaluate electrode/arc position for optimum scrap melting and 3) establish reduced order model for CFD data-base for fast model calculation.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Enhanced Lighting Signals for Safety and Efficiency - Experiments With Addressable LEDs

For enhanced roadway safety, clear and immediate visual cues are essential for preventing accidents between drivers and pedestrians. At intersections, however, the line of sight to other roadway users may be obstructed by vehicles and infrastructure. Additionally, adverse conditions, including low visibility, poor weather, or inadequate lighting can increase the potential for collisions. Distracted drivers and pedestrians can further exacerbate the risk of accidents, particularly when using a smartphone, rather than focusing on roadway surroundings. These issues demonstrate the need for infrastructure upgrades that enhance visibility and awareness at crosswalks. A potential solution is through enhanced lighting signals integrated into the roadway infrastructure. One such example is the use of addressable LEDs, individually controllable lights that can change color and brightness instantaneously through programmable microcontrollers. They can be installed and integrated into crosswalks to maintain visibility in conditions where pedestrians may be difficult to see, while also offering peripheral cues to pedestrians who may be distracted by their phones or other objects rather than the road. Such a system (as one example) that is integrated into a traffic intersection digital twin that tracks all roadway users accurately, has the potential to enhance visibility of vulnerable road users, and thus enhance safety. This paper examines the potential implementation and feasibility of this technology, as well as the safety benefits it could provide. Laboratory experiments with addressable LEDs reveal the capabilities and challenges of this technology for roadway infrastructure safety. These findings could pave the way for more integrated lighting in infrastructure for vehicles and pedestrians at intersections, merge and diverge locations, and other areas where complex interactions present safety hazards. Such lighting solutions, enabled by modern computation and communications, could enhance safety and efficiency in our transportation system and improve overall mobility.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Towards Sustainable Post-Exascale Leadership Computing

As computing systems approach the limits of traditional silicon technology, the diminishing returns in performance per watt present a significant barrier to sustaining growth in HPC. From a large-scale scientific supercomputing facility point of view, we propose a multifaceted strategy toward specialized hardware and architectures that are optimized for energy efficiency in specific applications. We also emphasize the need for integrating energy-aware practices across all levels of HPC, from system design and software development to operational policies. We discuss strategic opportunities such as the adoption of application-specific accelerators, the development of energy-efficient algorithms, and the implementation of data-driven operational analytics. Our goal is to develop a comprehensive roadmap ensuring that future leadership systems at OLCF can meet scientific demands while operating within stringent energy budgets, thereby supporting sustainable computing growth.

Shin, Woong↗