Search NASA⌕ Search

DOE OSTI · 1725776

Using machine learning to derive cloud condensation nuclei number concentrations from commonly available measurements

Abstract

Cloud condensation nuclei (CCN) number concentrations are an important aspect of aerosol–cloud interactions and the subsequent climate effects; however, their measurements are very limited. We use a machine learning tool, random decision forests, to develop a random forest regression model (RFRM) to derive CCN at 0.4 % supersaturation ([CCN0.4]) from commonly available measurements. The RFRM is trained on the long-term simulations in a global size-resolved particle microphysics model. Using atmospheric state and composition variables as predictors, through associations of their variabilities, the RFRM is able to learn the underlying dependence of [CCN0.4] on these predictors, which are as follows: eight fractions of PM 2.5 (NH 4 , SO 4 , NO 3, secondary organic aerosol (SOA), black carbon (BC), primary organic carbon (POC), dust, and salt), seven gaseous species (NO x , NH 3 , O 3 , SO 2 , OH, isoprene, and monoterpene), and four meteorological variables (temperature (T), relative humidity (RH), precipitation, and solar radiation). The RFRM is highly robust: it has a median mean fractional bias (MFB) of 4.4 % with ≈96.33 % of the derived [CCN0.4] within a good agreement range of -60% 2.5 speciation (NH 4 , SO 4 , NO 3 , and organic carbon (OC)), NO x , O 3 , SO 2 , T, and RH, as well as [CCN0.4] are available. We modify, optimize, and retrain the developed RFRM to make predictions from 19 to 9 of these available predictors. This retrained RFRM (RFRM-ShortVars) shows a reduction in performance due to the unavailability and sparsity of measurements (predictors); it captures the [CCN0.4] variability and magnitude at SGP with ≈67.02 % of the derived values in the good agreement range. This work shows the potential of using the more commonly available measurements of PM 2.5 speciation to alleviate the sparsity of CCN number concentrations' measurements.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Nair, Arshad Arjunan, Yu, Fangqun. 2020-11-05. Using machine learning to derive cloud condensation nuclei number concentrations from commonly available measurements. https://doi.org/10.5194/acp-20-12853-2020

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related reports

Water4Energy Step-1 Band-M Ready-to-Train Samples for TVA Weeks-to-Years Prediction, Version 0

AI-ready Band-M (monthly) labelled training pack for the Water4Energy Genesis Task-1 project on weeks-to-years prediction of Tennessee Valley temperature and precipitation. The deposit includes leakage-aware issue-time samples (samples_M_v0.nc; N=486), train-only scalers, issue-time split table, supporting monthly panels, and Python generation scripts to recreate the pack from the companion Tier-1 raw observation collection (https://doi.org/10.13139/ORNLNCCS/3398576). Each sample pairs a 12-month lookback of teleconnection indices and SST box anomalies with TVA-mean ERA5 anomaly targets (t2m, tp, msl) at leads 1–3 months.

54 ENVIRONMENTAL SCIENCES↗

Multi-Angle Snowflake Camera, particle analysis

The c1 level data product for the Mutli-Angle Snowflake Camera contains snowflake fall speeds and particle size, among other analysis for images associated with each hydrometeor.

54 ENVIRONMENTAL SCIENCES↗

Multi-Angle Snowflake Camera, time bins

The c1 level data product for the Mutli-Angle Snowflake Camera contains snowflake fall speeds and particle size, among other analysis for images associated with each hydrometeor.

54 ENVIRONMENTAL SCIENCES↗