Search NASASearch

SEARCH · Search NASA

Results for “data center metrics”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Advanced Data Center Energy Opportunities: Cloud and Infrastructure CoP - Data Center Energy and Efficiency with AI Adoption

The NLR portion of the "Cloud & Infrastructure CoP - Data Center Energy and Efficiency with AI Adoption" web meeting will cover data center locations, energy use and load growth, best practices, performance metrics, transition to direct liquid cooled data center equipment, and NLR's approach to optimizing data center.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI

Best Practices Guide for Energy-Efficient Data Center Design

This guide provides an overview of best practices for energy-efficient data center design which spans the categories of information technology (IT) systems and their environmental conditions, data center air management, cooling and electrical systems, and heat recovery. IT system energy efficiency and environmental conditions are presented first because measures taken in these areas have a cascading effect of secondary energy savings for the mechanical and electrical systems. This guide concludes with a section on metrics and benchmarking values by which a data center and its systems energy efficiency can be evaluated. No design guide can offer “the most energy-efficient” data center design but the guidelines that follow offer suggestions that provide efficiency benefits for a wide variety of data center scenarios.

29 ENERGY PLANNING, POLICY, AND ECONOMY

Modeling Framework for Data Center

This chapter highlights the critical need for advanced modeling of data centers due to their rapidly increasing energy consumption and impact on grid reliability. Driven by the demand for AI applications, data centers are projected to consume a significant portion of US energy by 2028, putting stress on an already challenged power grid. The chapter emphasizes the importance of "fast" time-scale models to understand the dynamic interactions between data centers and the grid, especially given the rapid power fluctuations of AI workloads. It outlines a modeling framework that includes both offline and real-time EMT domain simulations, detailing the necessary representations for various components like utility interfaces, transformers, IT loads, UPS, cooling loads, Battery Energy Storage Systems (BESS), generators, protection systems, and higher-level control systems. While standard simulation tools like PSCAD offer basic models, custom development is often required to accurately capture the unique and fast-changing behaviors of modern data centers. The chapter also discusses key metrics and test cases for validating these models, focusing on transient load responses, protection relay coordination, and demand flexibility. Finally, it addresses the challenges of modeling large-scale data centers, such as computational complexity and the trade-off between model fidelity and practicality, suggesting hybrid modeling approaches as a solution. The overarching goal is to create a robust framework that helps assess data center impacts on grid stability, identify vulnerabilities, and inform the development of standards for reliable integration of these large loads into the bulk power system.

25 ENERGY STORAGE

Artificial Intelligence for Data Center Operations (AIOps): Cooperative Research and Development (Final Report)

High performance computing data centers will increasingly need to rely on automation to keep pace with exascale growth in compute capability and to manage and optimize the data center environment and facility resources. Artificial intelligence and machine learning approaches provide the means to improve HPC data center operational efficiency, by learning historical trends and training models to operate on real-time data collected from both IT and facilities sources. NREL has developed methods of real-time collection, aggregation and streaming of these data in the ESIF HPC Data Center and has collected a significant dataset of relevant metrics across computer systems, racks, environmental, building and utility sources for research into various predictive analytics problems. HPE's Advanced Technology Group (ATG) is doing comprehensive research into exascale monitoring and management for High Performance Computing (HPC) systems (hereinafter HPE's Data Monitoring/ Management Technology). NREL and HPE will collaborate to add Artificial Intelligence (AI) to NREL's real-time data collection/ aggregation/ streaming system and HPE's Data Monitoring/ Management System, with the goal of improving the operational efficiency of NREL's Energy Systems Integration Facility (ESIF) HPC Data Center through data analytics on both historical and real-time data from IT systems and facilities operations. This collaboration will consist of efforts in Data Management, Data Analytics, and AI/ML Optimization for both manual and autonomous intervention in data center operations. This will be a multi-year, multi-staged effort with a goal towards building capabilities for an Advanced Smart Facility, and demonstration of these techniques in the NREL ESIF HPC Data Center.

97 MATHEMATICS AND COMPUTING

NLR HPC Facility Power Usage Effectiveness (PUE) Data

Timeseries of Energy Systems Integration Facility (ESIF) Data Center Power Usage Effectiveness (PUE) Data provided in Parquet and compressed CSV formats Power Metrics Timeseries Fields: ts: Timestamp cooling_kw: Cooling (kilowatts) - Captures the power used by fans and pipe trace heaters associated with outdoor cooling equipment. The dedicated tower filter pump power is also captured as cooling load. energy_reuse: Energy Reuse Effectiveness hvac_kw: Heating, ventilation, and air conditioning (kilowatts) - Captures fan walls, fan coils that support the data center electrical rooms, and the make-up air unit. it_power_kw: IT equipment (kilowatts) - Captures power used by the IT equipment on the data center floor. plug_and_light_kw: Lights and utility plugs (kilowatts) - Captures power associated with the data center and dedicated mechanical room. The crank-case heater for the emergency standby generator is also captured as light and plug load. pue: Power Usage Effectiveness pump_kw: Pumps (kilowatts) - Captures power from pumps that move water in the data center Energy Recover Water loop and the Tower Water loops, and also captures power used by the boost pumps that circulate water through the fan walls. Note: The tower filter pump runs constantly to filter water from the data center cooling tower system, so 2.67 kilowatts are attributed to this pump and that is not reflected in this data field. day: Day of month Outside Weather Station Timeseries Fields: ts: Timestamp outside_air_humidity: Outside air humidity - Relative humidity percent outside_air_temp: Outside air temperature - Degrees Fahrenheit day: Day of month More detail: High-Performance Computing Data Center Power Usage Effectiveness

97 MATHEMATICS AND COMPUTING

Expert and operator perspectives on barriers to energy efficiency in data centers

Abstract It was last estimated in 2016 that data centers (DCs) comprise approximately 2% of total US electricity consumption. However, this estimate is currently being updated to account for the massive increase in computing needs due to streaming, cryptocurrency, and artificial intelligence (AI). To prevent energy consumption that tracks with increasing computing needs, it is imperative we identify energy efficiency strategies and investments beyond the low-hanging fruit solutions. In a two-phased research approach, we ask: What non-technical barriers still impede energy efficiency (EE) practices and investments in the data center sector, and what can be done to overcome these barriers? In particular, we are focused on social and organizational barriers to EE. In Phase I, we performed a literature review and found that technical solutions are abundant in the literature, but fail to address the top-down cultural shifts that need to take place in order to adapt new energy efficiency strategies. In Phase II, reported here, we interviewed 16 data center operators/experts to ground-truth our literature findings. Our interview protocols focus on three aspects of DC decision-making: procurement practices, metrics and monitoring, and perceived barriers to energy efficiency. We find that vendors are the key drivers of procurement decisions, advanced efficiency metrics are facility-specific, and there is convergence in the design of advanced facilities due to the heat density of parallelized infrastructure. Our ultimate goals for our research are to design DC decarbonization policies that target organizational structure, empower individual staff, and foster a supportive external market.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI

A Methodology to Evaluate the Grid Reliability Impact of Oscillations Induced by Large Loads

The rapid growth of hyperscale AI data centers is bringing renewed attention to the reliability risk that sustained forced oscillations pose to bulk power systems, with cyclic computational workloads emerging as a new forcing source. Unlike the broadband, stochastic disturbances from traditional industrial loads such as arc furnaces, AI training and inference facilities can inject large active power swings concentrated at specific frequencies over extended durations - characteristics that existing grid planning practices do not account for. While the North American Electric Reliability Corporation (NERC) has recognized this gap and called for system-level studies of large load interconnections, no standardized methodology exists to screen, simulate, and quantify these risks at the planning stage. This report presents the Risk Assessment Tool for Large Load-induced Events (RATLLE), a Python-based, publicly available script suite developed at the Pacific Northwest National Laboratory to evaluate bulk power system reliability risks from data center-induced oscillations. RATLLE implements a three-module workflow: a screening module that identifies vulnerable interconnection locations and excitable system modes; a simulation module that models cyclic data center load behavior using a commercial positive sequence simulation platform; and an analysis module that computes risk metrics and generates interactive visualization dashboards. The risk metrics, formulated around simulation observables, map oscillation impacts to a three-stage severity scale spanning latent equipment fatigue through imminent cascading failure. The methodology is demonstrated on two Western Electricity Coordinating Council (WECC) system models: a publicly available 240-bus reduced representation and a detailed 2031 Heavy Winter planning case. Case studies illustrate that even modest 50 MW forced oscillations at resonant frequencies can produce wide-area power swings, N-1 security constraint violations, and cascading generator trips through protection actions - outcomes that would not occur under normal operating conditions without oscillations present. The results underscore the need for standardized oscillation impact assessment in large load interconnection studies and provide a reproducible, extensible framework for utilities to adopt or customize within their existing planning workflows.

Biswas, Shuchismita

Resilient Communities, Maryland (RCM): A Framework for Community-Driven Energy Resilience (Final Technical Report)

The Resilient Communities, Maryland (RCM) project integrates community-based participatory research (CBPR) approaches into energy resilience planning. The combination of qualitative, community-driven data and quantitative utility data improves both the effectiveness and efficiency of assessing the impacts of disruptions to energy infrastructure. Centered around a metric of critical services access, RCM created a repeatable framework for evaluating and modeling energy resilience in communities and promoted community engagement, resulting in more equitable stakeholder participation and improved decision-making processes for siting infrastructure to improve community energy resilience.

14 SOLAR ENERGY

WE-Validate: An Open-Source Framework For Wind Power Validation

Grid operators rely on historical weather time series at existing and planned wind power plants to make informed decisions when planning for a future power grid with very high penetration of renewable power. While synthetic wind power time series have been developed based on historical weather models, their validation with actual power production data remains complex due to variations in modeling practices and methodologies. This paper introduces the WE-Validate framework, originally designed for wind speed validation and now enhanced for wind power validation with a graphical user interface to support users with minimal programming experience. Validation of wind power with WE-Validate is based on robust metrics consisting of RMSE, centered RMSE, average bias, average percent bias, mean absolute error, mean absolute percent error, cross correlation, and calculation of ramping magnitude, rate, and duration. This paper showcases WE-Validate with validation of synthetically derived power for a wind plant in Washington state for one month in 2018. Validation of the synthetic power from two comparison data sets compared with observations shows both comparison series have strong correlation with observed across weekly and monthly aggregations while suffering from persistent negative bias. The suite of metrics within WE-Validate facilitates immediate insight into the utility of the comparison data sets through compression across multiple axes. This user-friendly, open-source tool can be extended beyond wind power, making it a valuable resource for system planners and operators in different domains.

Moncheur de Rieudotte, Malcolm P.

District heating utilizing waste heat of a data center: High-temperature heat pumps

Data centers are energy-intensive facilities with substantial low-grade waste heat. High-temperature heat pumps can be critical in boosting the data center’s waste heat for district heating, improving the system-level energy efficiency of data centers, and reducing CO 2 emissions in district heating. This study built thermodynamic models to assess high-temperature heat pumps with six configurations using low global warming potential refrigerants to supply heat up to 120 °C. The heat pump configurations include single-stage or two-stage cycles with advanced components, such as internal heat exchanger, economizer, flash tank, or parallel compressor. The refrigerants include R1234ze(Z), R1233ed(E), R1224yd(Z), R600, and R600a, and R245fa is used as a reference. A case study was carried out to recover the waste heat from the Frontier high-performance computing data center and provide hot water for district heating at the US Department of Energy’s Oak Ridge National Laboratory campus. The optimized performance of high-temperature heat pumps is characterized with various effectiveness of internal heat exchangers, and the operating parameters of economizer or flash tank, as well as their combination. The results show that the configurations of two-stage cycles with internal heat exchanger + flash tank and internal heat exchanger + economizer/parallel-compressor provide the highest coefficient of performance under scenarios of the maximum allowable value and a fixed value (0.3) of the internal heat exchangers’ effectiveness, respectively. R1234ze(Z) and R600a are the most promising refrigerants, considering trade-offs between the coefficient of performance and the volumetric heating capacity. The single-stage cycle with internal heat exchanger + economizer/parallel-compressor using R1234ze(Z) is recommended for utilizing Fronter’s waste heat in district heating. A one mega-watt high-temperature heat pump will reduce 33,100–33,200 metric tons of CO2 emission annually, corresponding to 85.4 %–85.6 % of equivalent CO2 emissions from natural gas boilers. Here, this study provides good guidelines for designing and deploying high-temperature heat pumps to support sustainable data centers and decarbonize district heating in the US.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI

Energy Systems Integration Facility Stewardship Summary: Fiscal Year 2025

A summary of NLR's stewardship of the nationally unique Energy Systems Integration Facility (ESIF) highlighting performance metrics, capability upgrades, and examples of R&D impact. 2025 brought a national focus on energy and ESIF is meeting the moment. All eyes are on data centers and domestic manufacturing and bringing the benefits of artificial intelligence to power system planning and operations. In step with national priorities, ESIF is building out capabilities that advance secure, reliable, and affordable power. ESIF hosted 190 multidisciplinary research projects, 855 high-performance computer users, and collaborated with 81 partners from industry, academia, research, and federal agencies. These research projects resulted in an AI method for detecting high-impedance faults with 90% accuracy, a grid controls demonstration in Connecticut, power quality validation of CorePower's flagship inductor, and a cybersecurity assessment of potential rogue capabilities in digitally connected energy devices. Facility infrastructure improvements enhanced the thermal research network, the SCADA system, the cyber range, power hardware-in-the-loop testing, and more. With support from the U.S Department of Energy (DOE), the ESIF laboratories continue to deliver leading solutions for secure, reliable, and affordable power.

24 POWER TRANSMISSION AND DISTRIBUTION

Evaluation of Saccadic Component Measure on Smooth Pursuit Tests

ABSTRACT Introduction Despite the advancement of eye-tracking technology for smooth pursuit (SP) eye movement evaluation, qualitative observation offers much information that is not captured by computers; hence, both objective and qualitative information should be utilized to evaluate SP. This study examined the consistency among our clinicians when evaluating SP using normal (N), grossly normal (GN), mildly abnormal (MA), and abnormal (AB) as classifications. We then evaluated the effect of combining GN and MA into a single subclinical (SUBC) category. We also evaluated the computerized percent saccade (PS) metric by determining its sensitivity and specificity in classifying SP. Materials and Methods Retrospective horizontal and vertical SP test videos and numerical data for 70 participants were obtained from the Neuro Kinetics Neuro-Otologic Test Center and de-identified. From this, eye-tracking videos, time plots of eye-tracking positional data, and tables of SP eye-tracking performance data were generated for 0.1, 0.3, and 0.5 Hz in both horizontal and vertical planes, totaling 6 tests per subject. Three clinicians rated each subject’s SP performance as N, GN, MA, or AB for a total of 6 ratings (3 frequencies, horizontal and vertical). This process was repeated using N, SUBC, and AB as rating categories. Clinicians also provided an overall SP rating for each plane as follows: AB if the results were abnormal for 2 or more frequencies tested. Alternatively, if fewer than 2 frequencies presented with a rating of AB, then an overall rating of MA, GN, or N was determined at the respective clinician’s discretion. Results When the 3 clinicians were tasked with classifying SP videos using 4 clinical categories, fair overall agreement was demonstrated. However, when MA and GN categories were combined into an SUBC category, the overall agreement for the 3 clinicians improved slightly for both horizontal SP (HSP) and vertical SP (VSP). This pattern of agreement did not differ considerably when comparing HSP versus VSP, and good consistency and reliability was observed across clinicians. Again, inter-rater consistency was smaller for VSP versus HSP despite the reduction in clinical categories. Cut-off values were generated for the PS metric and demonstrated good specificity and sensitivity when they were exceeded for 2 or more frequencies in a particular plane when evaluating a subject’s SP test. Conclusions

General & Internal Medicine

Systematic Benchmarking of Climate Models: Methodologies, Applications, and New Directions

As climate models become increasingly complex, there is a growing need to comprehensively and systematically assess model performance with respect to observations. Given the increasing number and diversity of climate model simulations in use, the community has moved beyond simple model intercomparison and toward developing methods capable of benchmarking a large number of simulations against a suite of climate metrics. Here, we present a detailed review of evaluation and benchmarking methods and approaches developed in the last decade, focusing primarily on scientific implications for Coupled Model Intercomparison Project (CMIP) simulations and CMIP6 results that contributed to the Intergovernmental Panel on Climate Change (IPCC) Sixth Assessment Report (AR6). Based on this review, we explain the resulting contemporary philosophy of model benchmarking, and provide clear distinctions and definitions of the terms model verification, process validation, evaluation, and benchmarking. While significant progress has been made in model development based on systematic evaluation and benchmarking efforts, some climate system biases still remain. The development of open‐source community software packages has played a fundamental role in identifying areas of significant model improvement and bias reduction. We review the key features of several software packages that have been commonly used over the past decade to evaluate and benchmark global and regional climate models. Additionally, we discuss best practices for the selection of evaluation and benchmarking metrics and for interpreting the obtained results, the importance of selecting suitable sources of reference data and accurate uncertainty quantification.

Environmental sciences

Engineering-Scale Test of a Water-Lean Solvent for Post-Combustion Capture

EPRI, Pacific Northwest National Laboratory, RTI International, and their project collaborators developed an engineering-scale test of a new water-lean solvent, N-(2-ethoxyethyl)-3-morpholinopropan-1-amine (EEMPA or 2-EEMPA) as a post-combustion CO 2 capture solvent for power plant applications. This test was conducted using the Pilot Solvent Test Unit at the National Carbon Capture Center. The primary objective of this test was to collect long-term data operating EEMPA with both coal- and natural gas-representative flue gases at the approximately 0.5 MW e -equivalent scale (5–10 metric tons CO 2 /day captured). This report details the activities preparing for that test, data collected during the test campaign, and analyses and interpretations of that data.

20 FOSSIL-FUELED POWER PLANTS

Data Center High-Temperature Liquid Cooling and Heat Reuse Techno-Economic Study: Preprint

Data centers are energy-intensive facilities with growing demands for efficiency and cost-effective operations. Smaller, more distributed edge inference data centers are expected to proliferate as AI applications require low latency closer to the user of AI tools, which presents a growing opportunity to explore the systems implications of liquid cooling on water and energy use. This study analyzes the implementation of high-temperature liquid cooling systems in a prototypical inference 1-MW data center and explores the potential for heat reuse across varying climates with a goal to optimize energy efficiency, reduce capital and operational costs, and identify opportunities for high-performance cooling and water use reduction infrastructure. This analysis evaluated configurations utilizing a peak day hourly sizing and systems performance spreadsheet to evaluate design and operational conditions from which component sizes, installed cost, operational cost, and performance metrics were determined for the Base case and the Elevated case. The techno-economic analysis included heat reuse applications across a range of heat recovery temperatures and heat rejection options. The analysis shows that high-temperature liquid cooling allows for improved energy efficiency, lower water consumption, and lower capital costs compared to traditional cooling approaches. Transitioning to elevated water inlet/outlet temperatures (50 degrees C/60 degrees C) eliminates the need for chillers, cooling towers, and heat recovery equipment in many scenarios across three distinct climate zones. This results in up to 75% capital cost savings for the cooling and heat recovery equipment, and with significantly reduced water consumption, especially in non-heat reuse applications. Heat generated from data centers can also be repurposed for space heating, domestic hot water, and other applications, and is most cost-effective when data center outlet temperatures exceed 55-60 degrees C.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI

Sensor Data Analytics and Data Quality Assessment Software

The proposed framework derives a set of quality metrics to provide critical insights into and tracking of grid operations, sensor performance, sensor longevity, and event statistics. Power grid engineers can utilize this information to identify problems with existing sensor locations and problematic power grid assets including generators, transmission lines, load centers, and substations. This information can also be used to identify unexpected/abnormal behavior of power grid components, improve power grid observability, and operational monitoring, and thus enhance real-time decision-making support system. Power grid planners can utilize this information to augment existing sensing architecture with new sensors and improve the observability of the network.

Mahapatra, Kaveri

Algal Biomass Production via Open Pond Algae Farm Cultivation: 2023 State of Technology and Future Research

The annual State of Technology (SOT) assessment is an essential activity for platform research conducted under the Bioenergy Technologies Office (BETO). It allows for the impact of research progress (both directly achieved in-house at the National Renewable Energy Laboratory [NREL] and furnished by partner organizations) to be quantified in terms of economic improvements in the overall biofuel production process for a particular biomass processing pathway, whether based on terrestrial or algal biomass feedstocks. As such, initial benchmarks can be established for currently demonstrated performance, and progress can be tracked toward out-year goals to ultimately demonstrate economically viable biofuel technologies. NREL's algae SOT benchmarking efforts historically focused both on front-end algal biomass production and separately on back-end conversion to fuels through NREL's "combined algae processing" (CAP) pathway. The production model is based on outdoor long-term cultivation data, enabled by comprehensive algal biomass production trials conducted under the Development of Integrated Screening, Cultivar Optimization, and Verification Research (DISCOVR) consortium efforts, driven by data furnished by Arizona State University (ASU) at the Arizona Center for Algae Technology and Innovation (AzCATI) testbed site. The CAP model is based on experimental efforts conducted primarily under NREL research and development projects. This report focuses on front-end algal biomass production, documenting the pertinent algal biomass cultivation parameters that were input to the NREL open pond algae farm model. Through partnerships under DISCOVR, collaborators at ASU furnished details on cultivation performance metrics including biomass productivity and harvest densities for recent growth trials done at the AzCATI site. The resulting biomass productivity was calculated at 16.7 g/m 2 /day (ash-free dry weight [AFDW], annual average) for seasonal cultivation of Picochlorum celeri TG2 and Monoraphidium minutum 26B-AM biomass strains at the ASU site. Picochlorum celeri achieved the best productivity from April to September, with Monoraphidium minutum 26B-AM being used between October and March. Tetraselmis striata LANL1001, usually part of the strain rotation in previous cultivation SOTs, was supplanted by Monoraphidium minutum 26B-AM in this year's outdoor cultivation trials. Finally, building from an industry case study presented in the 2022 SOT report, in the Appendix of this report we provide an update on further improved data furnished by an industry collaborator and resultant impacts on economics reflecting several seasonal scenarios. This case study provides a supplementary datapoint on work being performed elsewhere with a more dedicated focus on improved compositional quality, producing biomass enriched in lipids as may be more optimal for conversion upgrading to fuels and products.

09 BIOMASS FUELS

Evaluating the potential of disaggregated memory systems for HPC applications

Summary Disaggregated memory is a promising approach that addresses the limitations of traditional memory architectures by enabling memory to be decoupled from compute nodes and shared across a data center. Cloud platforms have deployed such systems to improve overall system memory utilization, but performance can vary across workloads. High‐performance computing (HPC) is crucial in scientific and engineering applications, where HPC machines also face the issue of underutilized memory. As a result, improving system memory utilization while understanding workload performance is essential for HPC operators. Therefore, learning the potential of a disaggregated memory system before deployment is a critical step. This paper proposes a methodology for exploring the design space of a disaggregated memory system. It incorporates key metrics that affect performance on disaggregated memory systems: memory capacity, local and remote memory access ratio, injection bandwidth, and bisection bandwidth, providing an intuitive approach to guide machine configurations based on technology trends and workload characteristics. We apply our methodology to analyze thirteen diverse workloads, including AI training, data analysis, genomics, protein, fusion, atomic nuclei, and traditional HPC bookends. Our methodology demonstrates the ability to comprehend the potential and pitfalls of a disaggregated memory system and provides motivation for machine configurations. Our results show that eleven of our thirteen applications can leverage injection bandwidth disaggregated memory without affecting performance, while one pays a rack bisection bandwidth penalty and two pay the system‐wide bisection bandwidth penalty. In addition, we also show that intra‐rack memory disaggregation would meet the application's memory requirement and provide enough remote memory bandwidth.

Ding, Nan