Search NASASearch

SEARCH · Search NASA

Results for “Cloud Computing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

How Cloud is Accelerating Research at NREL

This presentation coincides with AWS's announcement of their new Parallel Computing Service (PCS) which allows for easy creation of HPC-style clusters in their AWS cloud computing platform. I helped them beta test this service before it was made generally available in August. AWS asked if we would be interested in discussing our experience with the PCS service, and our experience with HPC workloads in the cloud in general, so this slideshow discusses a brief history of scientific computing at NREL and shares a bit of our experiences and approach to utilizing cloud services for HPC-style workloads.

97 MATHEMATICS AND COMPUTING

If We Build Them, They Will Run: Automated HPC Apps Deployment and Profiling with eBPF in Cloud

The high performance computing (HPC) community is in a period of transition. The rise of AI/ML coupled with a changing landscape of resources deems portability a new metric of performance, and methods to move between on-premises and cloud environments and assess compatibility are paramount. Here we design and test a strategy for bridging the gap between traditional HPC and Kubernetes environments – first containerizing applications, providing automated orchestration to run studies, and packaging the setup with automated means to assess performance using low overhead eXtended Berkeley Packet Filter (eBPF) programs. We first assess different designs for eBPF collection, demonstrating a tradeoff between number of programs deployed on a node and overhead added. We develop 5 low overhead eBPF programs that combine with streaming ML models to assess CPU, futex, TCP, shared memory, and file access across four different builds of an HPC application for CPU and GPU. We use eBPF data to generate insights into the possible underlying etiology of scaling issues. We then assess compatibility of a well-known benchmark, HPCG, across matrices of micro-architectures and optimization levels (217 containers across 24 instance types and over 7500 runs). We provide to the community 30 applications to deploy in our automated setup and perform a scaling study from 4 to a maximum of 256 nodes for both CPU and GPU applications. Finally, we use our gained knowledge about performance to generate compatibility artifacts that are used by a newly developed Kubernetes controller to intelligently select instance type based on optimizing a figure of merit. Along with insights to scaling in this environment with a collection of applications and templates to work from, we provide an overall strategy for approaching HPC application deployment and image selection based on compatibility in cloud.

Computer science

A Concept of a Convection–Cloud Chamber to Study Aerosol–Cloud–Drizzle Interactions

Understanding and quantifying the full chain of processes from aerosol activation to drizzle formation, and the associated feedbacks to the aerosol chemical and physical properties, all within a turbulent cloud are some of the toughest challenges in atmospheric chemistry and physics and are keys to the cloud–precipitation puzzle. This paper describes a concept for a new type of research facility consisting of a cloud chamber plus associated instrumentation and computational models, to explore aerosol–cloud interactions and processing, cloud optical properties, entrainment–cloud interactions, and quantitative assessment of drizzle onset. The envisioned design is for a 3 m × 3 m × 9 m chamber, such that the height is sufficient to achieve long lifetimes for aerosol processing and for significant drizzle growth by collision and coalescence. A suite of computational tools for simulating microphysical properties in the chamber provides a digital twin for designing the chamber and a range of example experiments. Theory and test results from novel remote sensing systems for exploring chemical and physical interactions and evolution of aerosols, cloud droplets, and drizzle within turbulent clouds are described. Testing of technology needed for the operation of a large-volume chamber, including aerosol generation methods and novel materials for water vapor boundary conditions, is described. Simulations suggest that spatially uniform turbulence and microphysical properties can be sustained in a steady state, with reasonable aerosol and water vapor fluxes, and that substantial drizzle can be produced through collision and coalescence of cloud droplets. Remaining challenges for more detailed engineering design and a discussion of possible first-light experiments are described.

54 ENVIRONMENTAL SCIENCES

GlideinBenchmark: collecting resource information to optimize provisioning

Choosing the right resource can speedup jobs completion, better utilize the available hardware and visibly reduce costs, especially when renting computers on the cloud. This was demonstrated in earlier studies on HEPCloud. But the benchmarking of the resources proved to be a laborious and time-consuming process. This paper presents GlideinBenchmark, a new Web application leveraging the pilot infrastructure of GlideinWMS to benchmark resources, and shows how to use the data collected and published by GlideinBenchmark to automate the optimal selection of resources. An experiment can select the benchmark or the set of benchmarks that most closely evaluate the performance of its workflows. With GlideinBenchmark and the help of the GldieinWMS Factory it controls the benchmark execution. Finally, a scheduler like HEPCloud’s Decision Engine can use the results to optimize resource provisioning.

Mambelli, Marco

GlideinBenchmark: collecting resource information to optimize provisioning

Choosing the right resource can speed up job completion, better utilize the available hardware, and visibly reduce costs, especially when renting computers in the cloud. This was demonstrated in earlier studies on HEPCloud. However, the benchmarking of the resources proved to be a laborious and time-consuming process. This paper presents GlideinBenchmark, a new Web application leveraging the pilot infrastructure of GlideinWMS to benchmark resources, and it shows how to use the data collected and published by GlideinBenchmark to automate the optimal selection of resources. An experiment can select the benchmark or the set of benchmarks that most closely evaluate the performance of its workflows. GlideinBenchmark, with the help of the GlideinWMS Factory, controls the benchmark execution. Finally, a scheduler like HEPCloud's Decision Engine can use the results to optimize resource provisioning.

Mambelli, Marco [Fermilab] (ORCID:0000000294892681

The remarkable inefficiency of stratocumulus

Marine stratocumulus clouds play a central role in Earth's climate system by reflecting incoming solar radiation and exerting a strong cooling effect. Their organization into open and closed mesoscale cellular morphologies can be thought of as an example of bistable dynamics driven by aerosol–cloud interactions and mesoscale processes. From the perspective of non-equilibrium thermodynamics, these structures are an example of a far-from-equilibrium open system that continuously produces and exports entropy. While entropy production has been studied in idealized deep convective systems, it has not yet been quantified for shallow clouds. Here, we compute and decompose the internal entropy production of open- and closed-cell stratocumulus using an ensemble of large-eddy simulations. We show that the overall entropy production of stratocumulus is low, reflecting the limited vertical extent and corresponding reduced ability to utilize the energy fluxes at the system's boundaries. Moist processes dominate the overall irreversibility, which, combined with their low entropy production, leads to a mechanical efficiency about an order of magnitude smaller than in deep convective systems. Although the dominant irreversible processes differ between open- and closed-cell regimes, the distributions of total entropy production largely overlap across the ensemble, limiting the ability to distinguish the dynamics of individual cases based solely on total entropy production.

54 ENVIRONMENTAL SCIENCES

torc (Torc Workflow Management System) [SWR-24-127]

This software package orchestrates execution of a workflow of jobs on distributed computing resources. It is optimized for use on HPCs with Slurm, but also can be used in the cloud and on local computers. Please refer to the documentation at https://nrel.github.io/torc

Thom, Daniel [National Renewable Energy Laboratory

Shifting institutional culture to develop climate solutions with Open Science

To address our climate emergency, “we must rapidly, radically reshape society”—Johnson & Wilkinson, All We Can Save. In science, reshaping requires formidable technical (cloud, coding, reproducibility) and cultural shifts (mindsets, hybrid collaboration, inclusion). We are a group of cross-government and academic scientists that are exploring better ways of working and not being too entrenched in our bureaucracies to do better science, support colleagues, and change the culture at our organizations. We share much-needed success stories and action for what we can all do to reshape science as part of the Open Science movement and 2023 Year of Open Science.

54 ENVIRONMENTAL SCIENCES

Shedding light on U.S. small and midsize data centers: Exploring insights from the CBECS survey

As demand for digital services accelerates, the energy and environmental footprint of data centers faces increasing scrutiny. While hyperscale cloud facilities have driven efficiency gains, small and midsize U.S. data centers remain a critical yet underexamined segment with significant untapped potential for energy savings. This study leverages data from the Commercial Buildings Energy Consumption Survey (CBECS) to analyze trends in server stocks, computing customers, cooling system adoption and efficiency, and geospatial distribution from 2012 to 2018. Findings reveal a sharp decline in small and midsize data centers, from 1.764 million to 1.398 million, with server counts dropping from 5.177 million to 4.262 million—aligning with the broader shift toward cloud computing. More than 40 % of servers in small data centers and 55 % in midsize data centers are housed in office buildings, and over half of all servers are concentrated in climate zones 5A (cold), 3A (mixed-humid), and 4A (mixed-humid), with the highest densities in metropolitan hubs. While direct expansion units remain the dominant cooling system, a clear transition toward more energy-efficient solutions, particularly air economizers, is evident. By integrating server and cooling system distributions, we estimate Power Usage Effectiveness (PUE) and Water Usage Effectiveness (WUE) for U.S. data centers by size and year. Results show that midsize data centers are more energy-efficient but more water-intensive due to the widespread use of water-cooled chillers. These findings highlight the trade-offs in cooling system selection and provide a critical foundation for policies aimed at enhancing efficiency in an evolving data center landscape.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI

Mesoscale Organization in Cumulus-Coupled Stratocumulus

Marine cloud systems cover a substantial portion of the world’s oceans. Most of these clouds form relatively close to the ocean surface, typically within one to two kilometers, a region referred to by meteorologists as the marine boundary layer. They are composed predominantly of liquid water, although ice particles can occur in mid- and high-latitude marine clouds during winter. In satellite imagery, these clouds appear bright against the darker ocean surface below, reflecting a large fraction of incoming sunlight back into space that would otherwise warm the ocean. Because marine boundary layer clouds cover such an extensive area of the ocean, they exert a significant influence on Earth’s overall transfer of solar energy absorbed by the surface and thermal energy emitted to space, a balance known as the planetary radiation budget. Marine boundary layer clouds are typically thin, and their formation and dissipation depend on a delicate balance between processes acting at the ocean surface below and the warm, dry air above. They are notoriously difficult to simulate accurately in weather forecast models, which often produce too few marine low clouds in midlatitudes and clouds in tropical regions that are excessively bright, meaning they reflect too much solar radiation. The marine boundary layer is frequently characterized by widespread overcast cloud cover that often transitions from a continuous, single-layer deck to more broken cloud fields toward the tropics. These transitions typically proceed through an intermediate stage in which shallow, broken clouds form beneath the overlying stratiform cloud deck. Once broken clouds develop below the overcast, they frequently self-organize into cloud clusters known as marine boundary layer convective complexes (MBLCCs), although the mechanisms governing the formation and organization of MBLCCs remain poorly understood. Accurately representing these transitions in long-range weather forecast models is essential because they influence the properties of air masses advected over the continental United States and Europe, and they become increasingly important for forecasts on seasonal and longer timescales. We employed two complementary approaches to investigate the processes controlling MBLCCs and their impact on marine cloud cover. Long-term observations from the U.S. Department of Energy’s Eastern North Atlantic (ENA) Observatory provided a unique dataset that allowed us to characterize fundamental properties of MBLCCs, including their typical size and frequency of occurrence. These observations were combined with high-resolution numerical simulations performed on supercomputers to examine the evolution of MBLCCs during cold-air outbreaks over the ENA region.

54 ENVIRONMENTAL SCIENCES

Classification of Cloud Particle Imagery and Thermodynamics (COCPIT): A New Databasing Tool for the Characterization of Cloud Particle Images Captured During DOE Field Campaigns

The Department of Energy for decades has explored the earth system and atmosphere through research and deployment of in-situ and remote sensing platforms during field campaigns. Among these datasets exists a vast supply of cloud particle images that provide visual insight into the complex microphysics in the clouds that span our globe. The millions of images collected over decades of deployments provides a unique opportunity to further our understanding of our atmosphere down to the crystal size. This work over the past 5 years has sought to organize these images into digestible datasets that can then be used by scientists to further our understanding of microphysics. A machine learning model was developed that categorizes over 1.5 million images across 11 weather events with over 90% accuracy according to particle type. The database was then extended to include dimensional characteristics of the particle as well as co-location of environmental properties, such as temperature and water content. Then, to initialize the connection between these data and our understanding of how crystals form and grow, weather research and forecasting simulations were run to generate the growth histories of the classified crystals. This research culminates with 2 databases per event: (1) a database of all classified crystals and their dimensional and environmental properties and (2) simulated growth histories of each crystal. Finally, a user interface was created to allow researchers to explore data statistics.

54 ENVIRONMENTAL SCIENCES

Efficient estimation of the modified Gromov–Hausdorff distance between unweighted graphs

Abstract Gromov–Hausdorff distances measure shape difference between the objects representable as compact metric spaces, e.g. point clouds, manifolds, or graphs. Computing any Gromov–Hausdorff distance is equivalent to solving an NP-hard optimization problem, deeming the notion impractical for applications. In this paper we propose a polynomial algorithm for estimating the so-called modified Gromov–Hausdorff (mGH) distance, a relaxation of the standard Gromov–Hausdorff (GH) distance with similar topological properties. We implement the algorithm for the case of compact metric spaces induced by unweighted graphs as part of Python library , and demonstrate its performance on real-world and synthetic networks. The algorithm finds the mGH distances exactly on most graphs with the scale-free property. We use the computed mGH distances to successfully detect outliers in real-world social and computer networks.

Oles, Vladyslav (ORCID:0000000188727463)

Workflow Provenance in the Computing Continuum for Responsible, Trustworthy, and Energy-Efficient AI

As Artificial Intelligence (AI) becomes more pervasive in our society, it is crucial to develop, deploy, and assess Responsible and Trustworthy AI (RTAI) models, i.e., those that consider not only accuracy but also other aspects, such as explainability, fairness, and energy efficiency. Workflow provenance data have historically enabled critical capabilities towards RTAI. Provenance data derivation paths contribute to responsible workflows through transparency in tracking artifacts and resource consumption. Provenance data are well-known for their trustworthiness helping explainability, reproducibility, and accountability. However, there are complex challenges to achieve RTAI, which are further complicated by the heterogeneous infrastructure in the computing continuum (Edge-Cloud-HPC) used to develop and deploy models. As a result, a significant research and development gap remains between workflow provenance data management and RTAI. In this paper, we present a vision of the pivotal role of workflow provenance in supporting RTAI and discuss related challenges. We present a schematic view between RTAI and provenance, and highlight open research directions.

Santos Souza, Renan

Machine learning reveals strong grid-scale dependence in the satellite N d –LWP relationship

The relationship between cloud droplet number concentration ( N d ) and liquid water path (LWP) is highly uncertain yet crucial for determining the impact of aerosol-cloud interactions (ACI) on Earth's radiation budget. The N d -LWP relationship is examined using a machine learning (ML) random forest model applied to five years of satellite data at grid resolutions ranging from 10° to 0.05° in 12 distinct regions. In the subtropics, the shape of the N d -LWP relationship switches from an inverted-V at 1° grid-resolution to an “M” shape at 0.1° resolution with decreased $\frac{\textrm{dln⁡LWP}}{\textrm{dln⁡}N_d}$ sensitivity. Tropical and midlatitude regions generally show a more positive sensitivity. Cloud sampling and filtering also influence this slope, wherein the exclusion of thin clouds, as commonly performed to reduce retrieval uncertainty, leads to strongly negative sensitivity across all regions. Precipitation is primarily responsible for driving the strength of the sensitivity, with strong positive slopes in raining clouds and negative and/or neutral responses found in non-raining clouds. A new method to compute radiative forcing from the ML model shows a robust Twomey radiative forcing across all regions and grid resolutions. However, LWP and cloud fraction adjustments to the radiative forcing, which are ∼50 % or smaller than the Twomey effect, decrease to negligible values with higher spatial resolution data. As Earth system models move toward higher spatial resolutions in the future, evaluating the LWP and CF adjustment contributions to the radiative forcing budget at these finer resolutions will be essential for evaluation and model development.

Aerosol-Cloud Interactions

Simulating Droplet-Resolved Haze and Cloud Chemistry Forming Secondary Organic Aerosols in Turbulent Conditions within Laboratory and Cloud Parcels

Most of our existing knowledge of cloud chemistry in regards to forming secondary organic aerosols (SOA) is based on measurements in bulk aqueous solutions. However, SOA reaction kinetics derived from bulk solution measurements might differ from the kinetics in actual cloud droplets, since turbulent mixing and ionic strengths, and ratio of surface area to volume in individual cloud droplets might vary substantially from bulk solutions in the real atmosphere. Three-dimensional models at various scales have been used to simulate aqueous chemistry. However, most of these models do not resolve turbulence down to the smallest length scales of 1 mm and do not simulate cloud chemistry in individual cloud droplets due to large computational costs. Here we incorporate the formation of isoprene epoxydiol SOA (IEPOX-SOA) in individual droplets within a one-dimensional explicit mixing parcel model (EMPM-Chem). We apply EMPM-Chem to simulate turbulence and droplet-resolved IEPOX-SOA formation using a configuration based on the Michigan Tech Pi chamber. We find that the dissolution of IEPOX gases is weighted more towards larger cloud droplets due to their large liquid water content (compared to smaller droplets), while the conversion of dissolved IEPOX to IEPOX-SOA is much greater within smaller deliquesced haze particles due to their higher acidity and ionic strengths compared to cloud droplets. We also find that as droplet residence times increase in the chamber, e.g., due to increasing ammonium bisulfate seed aerosol injection rates and/or increasing heights of the chamber, formation of IEPOX-SOA increases substantially. Thus, our EMPM-Chem model could be used to design future cloud chambers to maximize SOA production from cloud chemistry. We also apply the EMPM-Chem model to simulate how IEPOX-SOA formation evolves in individual cloud droplets within rising cloudy parcels in the atmosphere. We find that as subsaturated air is entrained into and turbulently mixed with the cloud parcel, evaporation causes a reduction in droplet sizes, which leads to corresponding increases in per droplet ionic strength and acidity. Increased droplet acidity in turn greatly accelerates the kinetics of IEPOX-SOA formation. In conclusion, our results provide key insights into single-cloud-droplet chemistry, suggesting that entrainment mixing may be an important process that increases SOA formation in the real atmosphere.

54 ENVIRONMENTAL SCIENCES

Portable Acceleration of CMS Computing Workflows with Coprocessors as a Service

Computing demands for large scientific experiments, such as the CMS experiment at the CERN LHC, will increase dramatically in the next decades. To complement the future performance increases of software running on central processing units (CPUs), explorations of coprocessor usage in data processing hold great potential and interest. Coprocessors are a class of computer processors that supplement CPUs, often improving the execution of certain functions due to architectural design choices. We explore the approach of Services for Optimized Network Inference on Coprocessors (SONIC) and study the deployment of this as-a-service approach in large-scale data processing. In the studies, we take a data processing workflow of the CMS experiment and run the main workflow on CPUs, while offloading several machine learning (ML) inference tasks onto either remote or local coprocessors, specifically graphics processing units (GPUs). With experiments performed at Google Cloud, the Purdue Tier-2 computing center, and combinations of the two, we demonstrate the acceleration of these ML algorithms individually on coprocessors and the corresponding throughput improvement for the entire workflow. This approach can be easily generalized to different types of coprocessors and deployed on local CPUs without decreasing the throughput performance. We emphasize that the SONIC approach enables high coprocessor usage and enables the portability to run workflows on different types of coprocessors.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND