Search NASA⌕ Search

SEARCH · Search NASA

Results for “representativeness”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 181 records · Page 10

Performance and Reliability Assessment of the U.S. Department of Energy Atmospheric Radiation Measurement (ARM) Data Advisor (ADA)

The Atmospheric Radiation Measurement (ARM) User Facility provides one of the world's largest openly accessible repositories of atmospheric observations through the ARM Data Discovery platform. Although the repository contains more than three decades of measurements collected from permanent observatories, mobile facilities, aircraft campaigns, and field experiments, identifying appropriate datasets can be challenging, particularly for new users unfamiliar with ARM instrumentation and datastream organization. To improve data accessibility, the ARM Data Center developed the ARM Data Advisor (ADA), an artificial intelligence-powered assistant designed to facilitate scientific data discovery, dataset interpretation, and user guidance. This report evaluates ADA's performance as a domain-specific scientific assistant using realistic atmospheric science workflows. The evaluation examines five key capabilities: data retrieval and curation efficiency, hallucination resistance, scientific reasoning, response to ambiguous queries, and content retention and session continuity. Representative prompts were developed to simulate typical interactions between researchers and the ARM Data Discovery platform, and ADA's responses were assessed for retrieval completeness, scientific accuracy, consistency, and practical usefulness. In these representative tests, ADA reduced the complexity of discovering and accessing ARM datasets by recommending appropriate datastreams, explaining instrumentation, interpreting metadata, and assisting with data processing workflows. ADA also exhibits strong domain knowledge of atmospheric science terminology and generally resists hallucination by acknowledging unavailable datasets and requesting clarification when appropriate. Overall, the results indicate that ADA represents a promising advancement in scientific data discovery within the ARM User Facility and has considerable potential to improve researcher productivity, particularly for new users and interdisciplinary scientists seeking efficient access to ARM observations.

Salvador, Christian [ORNL] (ORCID:0000000283287777↗

A Tensor Network-Based Quantum Algorithm for the Nonlinear 1D Burgers' Equation

In this work, we implement a tensor network-based quantum algorithm to solve unsteady, nonlinear partial differential equations (PDEs). The challenge lies in how to effectively represent, encode, process, and evolve the nonlinear system of PDEs on quantum computers. We will discuss the new techniques using the compressible 1-dimensional (1D) Burgers' equation as an example, because it represents the fundamental nonlinear feature and yet removes certain complexity in physics, allowing us to focus on the design of quantum algorithms. Previous attempts to solve nonlinear PDEs in quantum computation have often involved storing multiple copies of solutions or employing linearizations. Neither is practical due to exponential scaling with evolution time or insufficient solution accuracy. Our framework is based on matrix product states (MPSs) and matrix product operators (MPOs). For example, the velocity field is represented by MPS, whereas the linear and nonlinear spatial differential terms of the velocity field are processed by MPOs. Our primary focus herein is to verify and validate the various tensor network components of the algorithm using solutions obtained by the classical algorithms on high performance computing (HPC) architectures. We use a classical time marching method to demonstrate the functionality of the tensor network operations to model the PDE and their robustness with the time evolution of the system. Our classical simulation results demonstrate the utility of tensor network-based operations in modeling nonlinear PDEs and highlight the necessity as well as potential advantages of using quantum simulations for these techniques.

Gopalakrishnan Meena, Murali [ORNL] (ORCID:0000000↗

UPGRADE-E: Understanding Patterns Guiding Residential Adoption and Decisions about Energy Efficiency

This work represents the largest and most comprehensive dataset to-date of responses surveyed from U.S. residents regarding home modifications and energy-related decision-making, including both objective measures describing the occupants and their homes and subjective measures describing the more unpredictable human factors behind residential decision-making. First, to develop the survey, we interviewed 121 individual decision-makers within their households regarding planned or completed projects. We used the insights of these interviews to design a survey that was distributed to 10,000 households across the U.S. The overarching topics approached in the survey include descriptive information about the respondent, their household, their home, home modifications they have made, and the human cognition-based contextual factors involved in home projects and decision-making, such as preferences, motivations, barriers, and information sources. The processed data from the survey was compiled into a dataset entitled UPGRADE-E: Understanding Patterns Guiding Residential Adoption and Decisions about Energy Efficiency and represents the basis for the analyses included in this paper. The dataset represents a rich repository of home energy technology and modification decision-making results, the largest of its kind to-date. The abundance of contextual considerations within this dataset provides a robust resource for continual analysis, with possibilities for considering cross-cuts of data from a variety of perspectives.

Fuentes, Tracy L↗

Investigation of Point-Contact Strategies for CFD Simulations of Pebble-Bed Reactor Cores

This study numerically investigated the effects of various contact strategies on the thermal hydraulic behavior within a structured bed of 100 explicitly modeled pebbles. Four contact strategies and two thermal hydraulic conditions were considered. The strategies to avoid contact singularities include decreasing the pebble diameter, increasing the pebble diameter, bridging the pebble surfaces near the contact region, and capping the pebble surfaces near the contact region. One strategy, Strategy 3a, which involves bridging with a cylinder equal to 10% of the pebble diameter, was selected as the baseline strategy because it addressed the contact singularity while minimizing the geometric changes that affect the bed porosity. The two thermal hydraulic conditions were full-power operation (Case 1) and pressurized loss of forced cooling or PLOFC (Case 2). Simulations of the conjugate heat transfer within the structured bed were performed using the Reynolds-averaged Navier–Stokes approach with the realizable k-ϵ turbulence model and two-layer all y + wall treatment. The thermal-fluid quantities of interest were compared between the contact strategies for each case. In Case 1, the hydraulic behavior was sensitive to the contact strategy, with large differences in the pressure drop (30%) and volume-average velocity (4%). The thermal behavior was not sensitive, with less than a 0.5% difference across the strategies. To better understand the separate effects of each heat transfer mode, Case 2 was divided into the following subcases: conduction (Case 2a); conduction/radiation (Case 2b); and conduction/radiation/convection (Case 2c). Case 2a represents an early phase of the PLOFC transient. Case 2b represents an intermediate phase of the PLOFC transient, with the pebble temperatures sufficiently high for the radiative heat transfer to be non-negligible. Case 2c represents a late phase of the PLOFC transient after the establishment of the natural circulation of the heat transfer fluid. For Case 2, large differences in the contact strategy were observed only in Case 2a with only conduction. The difference in the maximum pebble temperature was 23% in Case 2a, 2% in Case 2b, and 0.3% in Case 2c.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Improving Photometric Redshift Estimates with Training Sample Augmentation

Abstract Large imaging surveys will rely on photometric redshifts (photo- z 's), which are typically estimated through machine-learning methods. Currently planned spectroscopic surveys will not be deep enough to produce a representative training sample for Legacy Survey of Space and Time (LSST), so we seek methods to improve the photo- z estimates that arise from nonrepresentative training samples. Spectroscopic training samples for photo- z 's are biased toward redder, brighter galaxies, which also tend to be at lower redshift than the typical galaxy observed by LSST, leading to poor photo- z estimates with outlier fractions nearly 4 times larger than for a representative training sample. In this Letter, we apply the concept of training sample augmentation, where we augment simulated nonrepresentative training samples with simulated galaxies possessing otherwise unrepresented features. When we select simulated galaxies with ( g - z ) color, i -band magnitude, and redshift outside the range of the original training sample, we are able to reduce the outlier fraction of the photo- z estimates for simulated LSST data by nearly 50% and the normalized median absolute deviation (NMAD) by 56%. When compared to a fully representative training sample, augmentation can recover nearly 70% of the degradation in the outlier fraction and 80% of the degradation in NMAD. Training sample augmentation is a simple and effective way to improve training samples for photo- z 's without requiring additional spectroscopic samples.

Moskowitz, Irene (ORCID:0000000222068589)↗

Field intercomparison of ice nucleation measurements: the Fifth International Workshop on Ice Nucleation Phase 3 (FIN-03)

Abstract. The third phase of the Fifth International Ice Nucleation Workshop (FIN-03) was conducted at the Storm Peak Laboratory in Steamboat Springs, Colorado, in September 2015 to facilitate the intercomparison of instruments measuring ice-nucleating particles (INPs) in the field. Instruments included two online and four offline measurement systems for INPs, which are a subset of those utilized in the laboratory study that comprised the second phase of FIN (FIN-02). The composition of the total aerosols was characterized using the Particle Analysis by Laser Mass Spectrometry (PALMS) and Wideband Integrated Bioaerosol Sensor (WIBS) instruments, and aerosol size distributions were measured by a laser aerosol spectrometer (LAS). The dominant total particle compositions present during FIN-03 were composed of sulfates, organic compounds, and nitrates, as well as particles derived from biomass burning. Mineral-dust-containing particles were ubiquitous throughout and represented 67 % of supermicron particles. Total WIBS fluorescing particle concentrations for particles with diameters of > 0.5 µm were 0.04 ± 0.02 cm−3 (0.1 cm−3 highest; 0.02 cm−3 lowest), typical of the warm season in this region and representing ≈ 9 % of all particles in this size range as a campaign average. The primary focus of FIN-03 was the measurement of INP concentrations via immersion freezing at temperatures > −33 °C. Additionally, some measurements were made in the deposition nucleation regime at these same temperatures, representing one of the first efforts to include both mechanisms within a field campaign. INP concentrations via immersion freezing agreed within factors ranging from nearly 1 to 5 times on average between matched (time and temperature) measurements, and disagreements only rarely exceeded 1 order of magnitude for sampling times coordinated to within 3 h. Comparisons were restricted to temperatures lower than −15 °C due to the limits of detection related to sample volumes and very low INP concentrations. Outliers of up to 2 orders of magnitude occurred between −25 and −18 °C; a better agreement was seen at higher and lower temperatures. Although the 5–10 factor agreement of INP measurements found in FIN-03 aligned with the results of the FIN-02 laboratory comparison phase, giving confidence in progress of this measurement field, this level of agreement still equates to temperature uncertainties of 3.5 to 5 °C that may not be sufficient for numerical cloud modeling applications that utilize INP information. INP activity in the immersion-freezing mode was generally found to be an order of magnitude or more, making it more efficient than in the deposition regime at 95 %–99 % water relative humidity, although this limited data set should be augmented in future efforts. To contextualize the study results, an assessment was made of the composition of INPs during the late-summer to early-fall period of this study inferred through comparison to existing ice nucleation parameterizations and through measurement of the influence of thermal and organic carbon digestion treatments on immersion-freezing ice nucleation activity. Consistent with other studies in continental regions, biological INPs dominated at temperatures of > −20 °C and sometimes colder, while arable dust-like or other organic-influenced INPs were inferred to dominate below −20 °C.

54 ENVIRONMENTAL SCIENCES↗

The need for carbon-emissions-driven climate projections in CMIP7

Abstract. Previous phases of the Coupled Model Intercomparison Project (CMIP) have primarily focused on simulations driven by atmospheric concentrations of greenhouse gases (GHGs), for both idealized model experiments and climate projections of different emissions scenarios. We argue that although this approach was practical to allow parallel development of Earth system model simulations and detailed socioeconomic futures, carbon cycle uncertainty as represented by diverse, process-resolving Earth system models (ESMs) is not manifested in the scenario outcomes, thus omitting a dominant source of uncertainty in meeting the Paris Agreement. Mitigation policy is defined in terms of human activity (including emissions), with strategies varying in their timing of net-zero emissions, the balance of mitigation effort between short-lived and long-lived climate forcers, their reliance on land use strategy, and the extent and timing of carbon removals. To explore the response to these drivers, ESMs need to explicitly represent complete cycles of major GHGs, including natural processes and anthropogenic influences. Carbon removal and sequestration strategies, which rely on proposed human management of natural systems, are currently calculated in integrated assessment models (IAMs) during scenario development with only the net carbon emissions passed to the ESM. However, proper accounting of the coupled system impacts of and feedback on such interventions requires explicit process representation in ESMs to build self-consistent physical representations of their potential effectiveness and risks under climate change. We propose that CMIP7 efforts prioritize simulations driven by CO2 emissions from fossil fuel use and projected deployment of carbon dioxide removal technologies, as well as land use and management, using the process resolution allowed by state-of-the-art ESMs to resolve carbon–climate feedbacks. Post-CMIP7 ambitions should aim to incorporate modeling of non-CO2 GHGs (in particular, sources and sinks of methane and nitrous oxide) and process-based representation of carbon removal options. These developments will allow three primary benefits: (1) resources to be allocated to policy-relevant climate projections and better real-time information related to the detectability and verification of emissions reductions and their relationship to expected near-term climate impacts, (2) scenario modeling of the range of possible future climate states including Earth system processes and feedbacks that are increasingly well-represented in ESMs, and (3) optimal utilization of the strengths of ESMs in the wider context of climate modeling infrastructure (which includes simple climate models, machine learning approaches and kilometer-scale climate models).

54 ENVIRONMENTAL SCIENCES↗

NeuralMie (v1.0): an aerosol optics emulator

The direct interactions of atmospheric aerosols with radiation significantly impact the Earth's climate and weather and are important to represent accurately in simulations of the atmosphere. This work introduces two contributions to enable a more accurate representation of aerosol optics in atmosphere models: (1) NeuralMie, a neural network Mie scattering emulator that can directly compute the bulk optical properties of a diverse range of aerosol populations and is appropriate for use in atmosphere simulations where aerosol optical properties are parameterized, and (2) TAMie, a fast Python-based Mie scattering code based on the Toon and Ackerman (1981) Mie scattering algorithm that can represent both homogeneous and coated particles. TAMie achieves speed and accuracy comparable to established Fortran Mie codes and is used to produce training data for NeuralMie. NeuralMie is highly flexible and can be used for a wide range of particle types, wavelengths, and mixing assumptions. It can represent core-shell scattering and, by directly estimating bulk optical properties, is more efficient than existing Mie code and Mie code emulators while incurring negligible error compared to existing aerosol optics parameterization schemes (0.08 % mean absolute percentage error).

54 ENVIRONMENTAL SCIENCES↗

Process-oriented evaluation of quasi-stationary Rossby waves and their impact on surface air temperature extremes in dynamical downscaling over North America

Quasi-stationary Rossby waves are a crucial component of the general circulation and play a significant role in regional water and energy cycles, as well as in extreme events. However, process-oriented evaluation for Rossby waves is rarely performed for dynamical downscaling simulations. To close this gap, we evaluate three classes of dynamical downscaling approaches, with a focus on quasi-stationary Rossby waves and their impact on surface air temperature over North America during Northern Hemisphere summer. The three classes of models differ in the way large-scale forcing is provided: a limited-area model (LAM) constrained only by lateral boundary conditions, represented by RegCM4 from the North American branch of the Coordinated Regional Downscaling Experiment (NA-CORDEX), a LAM with spectral nudging to maintain consistency in large-scale dynamics with the forcing data, represented by the Weather Research and Forecasting (WRF) model simulation in NA-CORDEX, and a global variable-resolution model with smoothly varying grid spacings, represented by the Community Atmosphere Model version 5.4, with the Model for Prediction Across Scales (MPAS) as its dynamical core (CAM-MPAS). With no constraints on the atmospheric dynamics, CAM-MPAS exhibits several mean biases in the upper-level circulations over the Pacific Coast region: a weaker subtropical jet, a northward-shifted mid-latitude jet, and an overestimated southerly flow. With the lateral boundary constraint alone, RegCM4 also exhibits weaker jets and overestimated southerly winds off the West Coast. Rossby ray theory reveals that those wind biases direct incoming Rossby waves northward. The erroneously routed Rossby waves distort the relationship between the accumulation of wave activity over the US West Coast and surface temperature anomalies over the Southern Great Plains, which emerges approximately 4 d after the convergence of wave-activity flux in the ERA-Interim reanalysis. Furthermore, the response of heatwaves to the extreme wave activity flux is not reproduced by the two models, a serious drawback as a dynamical downscaling framework is expected to connect large-scale forcing to local-scale phenomena. The WRF model employing spectral nudging is largely free from the aforementioned problems. A pair of sensitivity simulations suggests that spectral nudging is the key to improving the dynamics of quasi-stationary Rossby waves and their impact on surface air temperature. Our results also demonstrate the effectiveness of Rossby wave diagnostics that allow for realistic background flows for assessing the credibility of dynamical downscaling over North America, where incoming Rossby waves propagate through complex circulation patterns before traveling across the continent.

54 ENVIRONMENTAL SCIENCES↗

Model-based analysis of solute transport and potential carbon mineralization in the active layer of a hillslope underlain by permafrost with seasonal variability and climate change

Permafrost carbon, stored in frozen organic matter across vast Arctic and sub-Arctic regions, represents a substantial and increasingly vulnerable carbon reservoir. As global temperatures rise, the accelerated thawing of permafrost releases greenhouse gases, exacerbating climate change. However, freshly thawed permafrost carbon may also experience lateral transport by groundwater flow to surface water recipients such as rivers and lakes, increasing the terrestrial-to-aquatic transfer of permafrost carbon. Mobilization and subsurface transport mechanisms are poorly understood and not accounted for in global climate models, leading to high uncertainties in the predictions of the permafrost carbon feedback. Here, we focus on a hillslope in Endalen Valley, Svalbard, as a representative example of a high-Arctic hillslope underlain by continuous permafrost. We analyze solute transport in the form of a non-reactive tracer representing dissolved organic carbon (DOC) using a physics-based numerical model with the objective to study governing cryotic and hydrodynamic transport mechanisms relevant for warming permafrost regions. We first analyze transport times for DOC pools at different locations within the active layer under present-day climatic conditions and proceed to study susceptibility for deeper ancient carbon release in the upper permafrost due to thaw under different warming scenarios. Results suggest that DOC in the active layer near the permafrost table experiences rapid lateral transport upon thaw due to saturated conditions and lateral flow, while DOC close to the ground surface experiences slower transport due to flow in unsaturated soil. Deeper permafrost carbon release exhibits vastly different transport behaviors depending on warming and thaw rate. Gradual warming leads to small fractions of DOC being mobilized every year, while the majority moves vertically through percolation and cryosuction. Abrupt thaw resulting from a single very warm year leads to faster lateral transport times, similar to active layer DOC released in saturated conditions. Lastly, we analyze the potential susceptibility of DOC to mineralization to CO 2 prior to export due to soil moisture and temperature conditions. We find that high liquid saturation during transport coincides with very low mineralization rates and potentially inhibits mineralization into CO 2 before export. Overall, the results highlight the importance of subsurface hydrologic and thermal conditions for the retention and lateral export of permafrost carbon by subsurface flow.

Hamm, Alexandra [Stockholm Univ. (Sweden)] (ORCID:↗

Performance of wind assessment datasets in United States coastal areas

The atmospheric dynamics that occur near the intersection of land and water offer exciting and challenging opportunities for wind energy deployment in coastal locations. New models and tools are continually being developed in support of wind resource assessment, and three recent products are explored in this work for their performance in representing characteristics of the wind resource at coastal locations: the Global Wind Atlas 3 (GWA3), the 2023 National Offshore Wind dataset (NOW-23), and the wind climate simulations that are a component of the Wind Integration National Dataset (WIND) Toolkit Long-Term Ensemble Dataset (WTK-LED Climate). These relatively new products are freely available and user-friendly so that anyone – from a utility-scale developer to a resident or business owner – can evaluate the potential for wind energy generation at their location of interest. The validations in this work provide guidance on the accuracy of wind resource assessments for coastal customers interested in installing small or midsize wind turbines (≤ 1 MW in capacity) to support energy needs at the residential, business, or community scale, such as the island and remotely located participants of the U.S. Department of Energy's Energy Transitions Initiative Partnership Project. At 23 coastal locations across the United States, dataset performance varies according to different evaluation metrics. All three recent datasets tend to overestimate the observed coastal wind resource. GWA3 produces the smallest annual average wind speed relative errors, whereas WTK-LED Climate is in best agreement in terms of representing diurnal wind speed cycles. NOW-23 is the highest performing of the datasets for representing seasonal and interannual trends in the coastal wind resource. While GWA3 and WTK-LED Climate are relatively insensitive to the dataset output heights selected for wind resource assessment at small and midsize wind turbine hub heights (20–60 m), significant variation in the NOW-23 representation of wind shear across the wind profile in the lowest 100 m of the atmosphere leads to notable differences in wind speed estimates according to the dataset output heights selected for evaluation. GWA3 exhibits challenges in the representation of observed wind speed diurnal cycles at small and midsize turbine hub heights, likely due to the dataset's consistent treatment of hourly wind speed trends regardless of altitude.

17 WIND ENERGY↗

IM3 Open Source Data Center Atlas

IM3 Open Source Data Center Atlas Description This dataset contains locations of existing data center facilities in the United States. Data center locations were derived from OpenStreetMap (OSM), a crowd-sourced database. Data points from OSM are processed in various ways to determine additional variables provided in the data including: facility area (square feet), associated US county, and US state. This dataset can be used to identify areas of concentrated data center development and inform government and private sector planning strategies for future buildout of data centers and the infrastructure necessary to support it. Usage Notes Validation of OSM-derived data center locations is an ongoing development under the IM3 project, and the database will be updated as new information becomes available. In some instances, both the data center area (e.g., campus) and individual data center buildings are included as overlapping areas in the database. Both values are retained. Data center points, buildings, and campus areas are provided as separate layers in the downloadable data package. Note that data items are not necessarily complete across layers. That is, a specific data center may only be present as a single point geometry in the "point" layer while other data centers are represented in both the campus and building layers. In some cases, data center campuses and/or buildings straddle a county boundary line. Mappings to both counties are retained in the database as separate rows. These data rows will have the same data center id information, but each will have different county information. Crowd-sourced data, by nature, relies on individuals and communities to provide information. As a result, some data may be missing where it has not yet been reported. As we collect information on additional data center locations and as OSM receives additional contributions, the database will be updated to capture additional data points not yet shown. Technical Information Data is available for download under the following formats: GeoPackage (GPKG) CSV Geospatial data is provided in the WGS84 (EPSG:4326) coordinate reference system. The GeoPackage download contains the following layers. See usage notes for more information. "point" "building" "campus" The "point" layer includes all data from OSM that had POINT geometry type (i.e., individual coordinates). The "building" layer includes all OSM data that did not have POINT geometry and where the building tag in the OSM export was neither equal to "no" or null. Data that did not meet the "point" or "building" qualification was assumed to be a facility campus and included in the "campus" layer. The dataset contains the following parameters. Variables provided by OSM are labeled with (OSM-provided). id - unique identification number (OSM-provided with prefix of "node/", "relation/" and similar attributes removed) state - name of US state state_abb - two letter US state abbreviation state_id - state ID number county - name of US county county_id - county ID number ref - reference numbers or codes (OSM-provided) operator - the name of the company, corporation, or person in charge facility (OSM-provided) name - name of facility (OSM-provided) sqft - surface area of facility polygon, measured in square feet. Only available for "building" and "campus" layers lat - latitude of data centroid point lon - longitude of data centroid point type – represented spatial information. One of "point", "building", or "campus". geometry – POLYGON geometry of area footprint (in "campus" and "building" layers) or POINT geometry of locations (in "point" layer). This parameter is not included in the csv download. Attribution Data center locations were derived from OpenStreetMap, which is made available at openstreetmap.org under the Open Database License (ODbL). US state and county boundary information was collected from the US Census Bureau for the year 2024, which is made publicly available at https://www.census.gov/geographies/mapping-files.html Acknowledgment IM3 is a multi-institutional effort led by Pacific Northwest National Laboratory and supported by the U.S. Department of Energy's Office of Science as part of research in MultiSector Dynamics, Earth and Environmental Systems Modeling Program. License The IM3 Open Source Data Center Atlas is made available under the Open Database License: http://opendatacommons.org/licenses/odbl/1.0/. Disclaimer This material was prepared as an account of work sponsored by an agency of the United States Government. Neither the United States Government nor the United States Department of Energy, nor the Contractor, nor any or their employees, nor any jurisdiction or organization that has cooperated in the development of these materials, makes any warranty, express or implied, or assumes any legal liability or responsibility for the accuracy, completeness, or usefulness or any information, apparatus, product, software, or process disclosed, or represents that its use would not infringe privately owned rights. Reference herein to any specific commercial product, process, or service by trade name, trademark, manufacturer, or otherwise does not necessarily constitute or imply its endorsement, recommendation, or favoring by the United States Government or any agency thereof, or Battelle Memorial Institute. The views and opinions of authors expressed herein do not necessarily state or reflect those of the United States Government or any agency thereof. PACIFIC NORTHWEST NATIONAL LABORATORYoperated byBATTELLEfor theUNITED STATES DEPARTMENT OF ENERGYunder Contract DE-AC05-76RL01830

Mongird, Kendall [Pacific Northwest National Labor↗

IM3 Open Source Data Center Atlas

IM3 Open Source Data Center Atlas Description This dataset contains locations of existing data center facilities in the United States. Data center locations were derived from OpenStreetMap (OSM), a crowd-sourced database. Data points from OSM are processed in various ways to determine additional variables provided in the data including: facility area (square feet), associated US county, and US state. This dataset can be used to identify areas of concentrated data center development and inform government and private sector planning strategies for future buildout of data centers and the infrastructure necessary to support it. Usage Notes Validation of OSM-derived data center locations is an ongoing development under the IM3 project, and the database will be updated as new information becomes available. In some instances, both the data center area (e.g., campus) and individual data center buildings are included as overlapping areas in the database. Both values are retained. Data center points, buildings, and campus areas are provided as separate layers in the downloadable data package. Note that data items are not necessarily complete across layers. That is, a specific data center may only be present as a single point geometry in the "point" layer while other data centers are represented in both the campus and building layers. In some cases, data center campuses and/or buildings straddle a county boundary line. Mappings to both counties are retained in the database as separate rows. These data rows will have the same data center id information, but each will have different county information. Crowd-sourced data, by nature, relies on individuals and communities to provide information. As a result, some data may be missing where it has not yet been reported. As we collect information on additional data center locations and as OSM receives additional contributions, the database will be updated to capture additional data points not yet shown. Data items will occasionally be removed from OSM if they are misidentified, if they no longer exist, if they are duplicates of another item, or similar. For that reason, updated versions of this database may not contain all data center locations included in previous versions. Technical Information Data is available for download under the following formats: GeoPackage (GPKG) CSV Geospatial data is provided in the WGS84 (EPSG:4326) coordinate reference system. The GeoPackage download contains the following layers. See usage notes for more information. "point" "building" "campus" The "point" layer includes all data from OSM that had POINT geometry type (i.e., individual coordinates). The "building" layer includes all OSM data that did not have POINT geometry and where the building tag in the OSM export was neither equal to "no" or null. Data that did not meet the "point" or "building" qualification was assumed to be a facility campus and included in the "campus" layer. The dataset contains the following parameters. Variables provided by OSM are labeled with (OSM-provided). id - unique identification number (OSM-provided with prefix of "node/", "relation/" and similar attributes removed) state - name of US state state_abb - two letter US state abbreviation state_id - state ID number county - name of US county county_id - county ID number ref - reference numbers or codes (OSM-provided) operator - the name of the company, corporation, or person in charge facility (OSM-provided) name - name of facility (OSM-provided) sqft - surface area of facility polygon, measured in square feet. Only available for "building" and "campus" layers lat - latitude of data centroid point lon - longitude of data centroid point type – represented spatial information. One of "point", "building", or "campus". geometry – POLYGON geometry of area footprint (in "campus" and "building" layers) or POINT geometry of locations (in "point" layer). This parameter is not included in the csv download. Attribution Data center locations were derived from OpenStreetMap, which is made available at openstreetmap.org under the Open Database License (ODbL). US state and county boundary information was collected from the US Census Bureau for the year 2024, which is made publicly available at https://www.census.gov/geographies/mapping-files.html Acknowledgment IM3 is a multi-institutional effort led by Pacific Northwest National Laboratory and supported by the U.S. Department of Energy's Office of Science as part of research in MultiSector Dynamics, Earth and Environmental Systems Modeling Program. License The IM3 Open Source Data Center Atlas is made available under the Open Database License: http://opendatacommons.org/licenses/odbl/1.0/. Disclaimer This material was prepared as an account of work sponsored by an agency of the United States Government. Neither the United States Government nor the United States Department of Energy, nor the Contractor, nor any or their employees, nor any jurisdiction or organization that has cooperated in the development of these materials, makes any warranty, express or implied, or assumes any legal liability or responsibility for the accuracy, completeness, or usefulness or any information, apparatus, product, software, or process disclosed, or represents that its use would not infringe privately owned rights. Reference herein to any specific commercial product, process, or service by trade name, trademark, manufacturer, or otherwise does not necessarily constitute or imply its endorsement, recommendation, or favoring by the United States Government or any agency thereof, or Battelle Memorial Institute. The views and opinions of authors expressed herein do not necessarily state or reflect those of the United States Government or any agency thereof. PACIFIC NORTHWEST NATIONAL LABORATORYoperated byBATTELLEfor theUNITED STATES DEPARTMENT OF ENERGYunder Contract DE-AC05-76RL01830

Mongird, Kendall [Pacific Northwest National Labor↗

A Generalized approach to the operationalization of Software Quality Models

Comprehensive measures of quality are a research imperative, yet the development of software quality models is a wicked problem. Definitive solutions do not exist and quality is subjective at its most abstract. Definitional measures of quality are contingent on a domain, and even within a domain, the choice of representative characteristics to decompose quality is subjective. Thus, the operationalization of quality models brings even more challenges. A promising approach to quality modeling is the use of hierarchies to represent characteristics, where lower levels of the hierarchy represent concepts closer to real-world observations. Building upon prior hierarchical modeling approaches, we developed the Platform for Investigative software Quality Understanding and Evaluation (PIQUE). PIQUE surmounts several quality modeling challenges because it allows modelers to instantiate abstract hierarchical models in any domain by leveraging organizational tools tailored to their specific contexts. Here, we introduce PIQUE; exemplify its utility with two practical use cases; address challenges associated with parameterizing a PIQUE model; and describe algorithmic techniques that tackle normalization, aggregation, and interpolation of measurements.

Data aggregation↗

Public Reference Data for Megawatt-Scale Hydrogen Electrolysis - Simulated Wave

The U.S. Department of Energy and the National Laboratory of the Rockies (NLR) demonstrate hydrogen electrolysis, hydrogen compression and storage, and variable hydrogen fuel cell power production using megawatt-scale equipment at NLR’s Flatirons Campus as part of the Advanced Research on Integrated Energy Systems (ARIES) initiative. This dataset represents part of that effort and is intended for academic, national laboratory, industrial, and other stakeholders to plan, design, and validate models of megawatt-scale hydrogen technologies and diverse energy infrastructure nationwide. These data provide a baseline for how existing hydrogen electrolysis technologies perform when coupled with various energy technologies. Future datasets will demonstrate how existing hydrogen fuel cell technologies can provide controllable, dispatchable, and variable power output for artificial intelligence (AI) data centers and other variable loads. This dataset entry describes hydrogen production using a single, simulated wave energy conversion device. The electrolyzer is a 1.25-MW proton exchange membrane type MC250 system manufactured by Nel Hydrogen. While the unit supports up to 2.5 MW of electrolysis, NLR only has a single 1.25-MW electrolysis stack. For the wave energy, NLR used a wave energy converter model from PacWave. These devices can be equipped with accumulators and pressure relief values to smooth the power output by storing and releasing hydraulic energy. Using a peak power output of 10 MW, the model created two 25-minute profiles: one with and one without the accumulators and pressure relief valves. To down select the profile data from the native resolution of 20 Hz to 1 Hz, NLR took the mean of every 20 data points. NLR experimented with two simulated wave energy power plants: one that peaks at 10 MW, and one that peaks at 5 MW. These profiles were scaled for the physical 1.25 MW electrolyzer by multiplying the original profiles by one eighth and one quarter, respectively. The first profile matches the capacity rating of eight of the 1.25 MW electrolyzers, while the second matches four electrolyzers. Finally, NLR experimented with two settings for the electrolyzer power supply minimum and maximum current ramp rates (gain and slew): 200 and 400 amperes per second. The simulated profiles were translated from power (kilowatts) to current (amperes) using a curve fit with calibration data and sent to the electrolyzer power supply at 1-Hz frequency. These datasets report relevant hydrogen balance-of-plant and system data, all captured at 1 Hz, including hydrogen mass production measured with an Emerson Coriolis flow meter. Each .zip file represents a single wave electrolysis experiment and is formatted as follows: {technology}-{accumulator?}_{number of 1.25 MW electrolyzers connected}-{electrolyzer ramp rate in amperes/second} For instance, “wavePacWave-Noacc_4-400.zip” represents the 25 minute-long experiment using the PacWave’s wave energy converter model, equipped with no accumulator, connected to four 1.25-MW electrolyzers with their power supplies set to a maximum current ramp rate (gain and slew) of 400 A/s. Each .zip folder contains the following files: A .csv file containing raw data. An .xlsx file explaining all the fields in the raw data. A .png plot showing the time series of hydrogen production in kilograms per hour, electrolysis power consumption, and input wave power. An experiment, labeled “characterization_200.zip”, demonstrates the MC250 electrolyzer steady-state response with 30 minute load steps for a total duration of 5 hours. Finally, a .csv file is provided with all wave profiles combined into one dataset labeled "combined_wave_experiments.csv". NLR also built an AI/machine-learning predictive model based on these datasets. The model ingests the electrolyzer current command in amperes, as well as various pressures and temperatures across the system, and predicts hydrogen output in kilograms per hour. The complete model can be found at https://huggingface.co/NatLabRockies/ptmelt-hydrogen-electrolysis.

08 HYDROGEN↗

Public Reference Data for Megawatt-Scale Hydrogen Electrolysis - Simulated Marine Hydrokinetic Tidal Turbine

The U.S. Department of Energy and National Laboratory of the Rockies (NLR) demonstrate hydrogen electrolysis, hydrogen compression and storage, and variable hydrogen fuel cell power production using megawatt-scale equipment at NLR’s Flatirons Campus as part of the Advanced Research on Integrated Energy Systems (ARIES) initiative. This dataset is part of that effort and is intended for academic, national laboratory, industrial, and other stakeholders to plan, design, and validate models of megawatt-scale hydrogen technologies and diverse energy infrastructure. These data provide a baseline for how existing hydrogen electrolysis technologies perform when coupled with other energy technologies. This dataset contains inputs and outputs from simulations of a floating marine hydrokinetic turbine over approximately half a tidal cycle (~6.6 hours). Inflow conditions were derived from field measurements in Alaska’s Cook Inlet and represent a tidal environment in which the current speed ramps from near 0 m/s to a peak of 3 m/s and back. The original acoustic doppler current profiler dataset is publicly available on the Marine and Hydrokinetic Data Repository. In a full tidal cycle, the flow reverses and the rotor would reorient; this reversal was not modeled. In the Cook Inlet campaign , turbulence intensity was similar in both directions. Two inflow cases are included. In the first case, labeled “raw” in the files, the measured current time series was used directly in the InflowWind module of OpenFAST. Speed and direction were applied as a function of time and elevation, uniformly in the horizontal direction. With full spatial coherence, this approach captures high turbulent variability and results in pronounced power fluctuations, so it is considered a conservative, near-worst-case representation of loading. In the second case, labeled “average” in the files, a 30-minute moving average was applied to extract the slowly varying mean speed. The residual fluctuations about this mean were used to generate spatially varying, full-field turbulence inputs with TurbSim, giving a more physically realistic representation of the inflow across the rotor disk. Two random realizations were used to produce distinct inflow conditions for two OpenFAST simulations representing a two-turbine array. The same turbulence intensity is applied across the full time series, producing larger fluctuations at the start and end, where the mean speed is low. The second case is the more appropriate framework for performance and power assessment but overpredicts turbulence at lower flow speeds and underpredicts it at higher speeds. As the floating platform moves and the rotor changes its x-position, Taylor’s frozen turbulence hypothesis used by InflowWind assumes a constant rather than a time-varying mean velocity, introducing some inaccuracy in the velocity plane sampling. The turbine modeled is the 500-kW Reference Model 1, a horizontal-axis two-bladed hydrokinetic turbine on a four-column floating semisubmersible substructure . Simulations were performed using OpenFAST v4.1 with the Reference Open Source Controller (ROSCO) v2.10. All input files required to reproduce the simulations are included. The electrolyzer is a 1.25-MW proton exchange membrane type MC250 system manufactured by Nel . This unit supports up to 2.5 MW, but NLR has only a single 1.25-MW stack. The datasets report hydrogen balance-of-plant and system data, all captured at 1 Hz, including hydrogen mass production measured with an Emerson Coriolis flow meter. The system controls hydrogen production by varying direct current applied to the stack, from a maximum of 3,000 A to a minimum safe operating current of 300 A, or 10%. Because the current–voltage characteristic changes as the stack ages and efficiency degrades, the actual minimum safe operating power changes over time. The simulated tidal turbine time series data was translated from power (kilowatts) to current (amperes) using a curve fit with calibration data and sent to the electrolyzer power supply at 1-Hz. Each zip file represents a single tidal electrolysis experiment and is named: {technology}_{inflow method}_{number of 500 kW tidal turbines connected} For instance, “tidal-500kW-RM1_average_2.zip” is a 6-hour experiment using the 500-kW tidal reference model, scaled by 2x (1-MW) to better match the electrolyzer maximum of 1.25MW, fed with the 30-minute moving average current case. Each zip folder contains the following files: A .csv file of raw data. An .xlsx file explaining all the fields in the raw data. A .png plot showing the time series of hydrogen production in kilograms per hour, electrolysis power consumption, and input wave power. A .csv file combines all tidal profiles as "combined_tidal_experiments.csv." A separate experiment, “characterization_200.zip,” shows the MC250 electrolyzer steady-state response with 30-minute load steps over 5 hours and is accessible with this entry.

08 HYDROGEN↗

In-Situ Species Concentration Measurements in Ammonia-Mix Flames Using FTIR Spectroscopy

Hydrogen and ammonia represent two carbon-free fuel sources that could be used in place of current fossil energy sources in combustion systems. To develop optimized ammonia combustion systems, validated modeling tools are needed. In the open literature, it has been shown that the complex chemistry associated with fuel-bound nitrogen contained in ammonia differs greatly from natural gas or hydrogen combustion. As a result, several new chemical kinetic mechanisms have been developed. Many of these mechanisms have been validated experimentally, however this has primarily focused on bulk parameters such as laminar flame speed and ignition delay time. Critically, high quality measurements of species concentrations are needed under controlled conditions which are easily represented by simple models. In this paper, direct, in-situ measurements of species concentrations and gas temperature are performed in a laminar flat-flame burner. This arrangement enables comparison with 1D model predictions, better isolating chemical kinetics from the fluid dynamics. Quantitative species concentrations are determined by absorption spectroscopy using an FTIR spectrometer. Fuel compositions representative of cracked ammonia (NH3/H2) and ammonia-natural gas (NH3/CH4) are considered for rich and lean equivalence ratios. A major focus of the paper is on the selection of spectral features for nitric oxide and ammonia and correcting for large amounts of baseline H2O absorption.

Bedick, Clinton↗

Bias Correction and Statistical Downscaling of Solar Radiation Using NA-CORDEX and the NSRDB

The current state-of-art for estimating long-term PV production uses long-term estimates of solar radiation variables, such as global horizontal irradiance (GHI), from previous years. This data is used in models such as the System Advisor Model (SAM) or PYSyst to predict annual production for a PV plant. This information is then used to estimate the production over the next 20 years (a typical plant lifetime) under the assumption that the variability over the current period is representative of the future. As the PV industry moves to extend plant lifetimes to 50 years the current assumptions of representativeness of weather may not be appropriate. This is especially true as our climate changes rapidly. To assess long-term PV production, future projections for solar radiation based on projected carbon emissions are readily available in regional and global climate models. However, climate model projections contain inherent biases that may need to be corrected for accurate analysis of future projections of climate variables. Several studies have analyzed projections of solar radiation for future years, however the accuracy of the model output compared to current and historic data has not been widely studied. Chen (2021) showed that available climate models do not accurately represent solar radiation in some cases, over-projecting GHI at the surface while under-projecting its obstructions, such as clouds and aerosols. This works aims to (1) increase understanding of the accuracy of solar radiation currently available in global and regional climate models and (2) implement bias correction through linear models based on reanalysis data compared to observed solar radiation. The latter aim will be conducted using available observed solar radiation data and modeled data from several regional climate models (RCMs). The bias correction method will be applied to projections of solar radiation resulting in a more accurate representation of the future of solar production.

climate data↗