Search NASASearch

SEARCH · Search NASA

Results for “ensemble data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 145 records · Page 8

An Overview of the NASA Spring/Summer 2008 Arctic Campaign - ARCTAS (Arctic Research of the Composition of the Troposphere from Aircraft and Satellites)

ARCTAS (Arctic Research of the Composition of the Troposphere from Aircraft and Satellites) is a major NASA led airborne field campaign being performed in the spring and summer of 2008 at high latitudes (http://cloud1.arc.nasa.gov/arctas/). ARCTAS is a part of the International Polar Year program and its activities are closely coordinated with multiple U. S. (NOAA, DOE), Canadian, and European partners. Observational data from an ensemble of aircraft, surface, and satellite sensors are closely integrated with models of atmospheric chemistry and transport in this experiment. Principal NASA airborne platforms include a DC-8 for detailed atmospheric composition studies, a P-3 that focuses on aerosols and radiation, and a B-200 that is dedicated to remote sensing of aerosols. Satellite validation is a central activity in all these platforms and is mainly focused on CALIPSO, Aura, and Aqua satellites. Major ARCTAS themes are: (1) Long-range transport of pollution to the Arctic including arctic haze, tropospheric ozone, and persistent pollutants such as mercury; (2) Boreal forest fires and their implications for atmospheric composition and climate; (3) Aerosol radiative forcing from arctic haze, boreal fires, surface-deposited black carbon, and other perturbations; and (4) Chemical processes with focus on ozone, aerosols, mercury, and halogens. The spring deployment (April) is presently underway and is targeting plumes of anthropogenic and biomass burning pollution and dust from Asia and North America, arctic haze, stratosphere-troposphere exchange, and ozone photochemistry involving HOx and halogen radicals. The summer deployment (July) will target boreal forest fires and summertime photochemistry. The ARCTAS mission is providing a critical link to enhance the value of NASA satellite observations for Earth science. In this talk we will discuss the implementation of this campaign and some preliminary results.

Jacob, Daniel J.

SWOT Applications in Alaska

Use of SWOT for model calibration and assimilation are encouraging, but further work is needed: account for uncertainty in meteorological forcing (calibration) and apply ensemble Kalman Smoother (data assimilation). Lessons learned can inform development of operational NOAA National Water Model Alaskan domain and other SWOT applications.

SWOT

Biological Research and Space Health Enabled by Machine Learning to Support Deep Space Missions

A key science goal of the NASA “Moon to Mars” campaign is to understand how biology responds to the Lunar, Martian, and deep space environments in order to advance fundamental knowledge, reduce risk, and support safe, productive human space missions. Through the powerful emerging computer science approaches of artificial intelligence (AI) and machine learning (ML), a paradigm shift has begun in biomedical science and engineered astronaut health systems, to enable Earth-independence and autonomy of mission operations. We present a decadal view of AI/ML architecture to support deep space mission goals, developed in concert with leaders in the field. We describe current AI/ML methods to support 1) fundamental biology, 2) in situ analytics, 3) high performance computing hardware, 4) automated science, 5) self-driving labs, 6) remote data management, 7) integrated real-time mission biomonitoring, and 8) a Precision Space Health system. Cutting-edge AI/ML approaches that can be integrated to support these domains include active learning, explainable AI, adaptive learning, causal inference, knowledge graphs, federated learning, transfer learning, and large language models. Finally, we present results from several current ML projects that are underway in the field to address key challenges of small sample n, high feature count, heterogeneity, and sparse data. These include 1) connecting omics data to phenotypic data using an ensemble model to infer causality of spaceflight rodent liver health disruption, 2) usage of explainable ML to interrogate the muscular underpinnings of spaceflight muscle atrophy, 3) ML models analyzing and determining directed acyclic graphs of human space health risk leveraging rodent bone datasets, 4) usage of large pre-trained models connecting biomedical knowledgebases with small spaceflight datasets to understand gene-to-gene interaction networks, and 5) a suite of benchmarked open science datasets (spaceflight mouse liver; radiation DNA damage) enabling programmers to identify the best ML algorithms to answer space biological science questions.

space biology

Biological Research and Space Health Enabled by Machine Learning to Support Deep Space Missions

A key science goal of the NASA “Moon to Mars” campaign is to understand how biology responds to the Lunar, Martian, and deep space environments in order to advance fundamental knowledge and support human space missions. Through artificial intelligence (AI) and machine learning (ML), a paradigm shift has begun in space biosciences and engineered astronaut health systems, to enable Earth-independence and mission operations autonomy. We describe current AI/ML methods to support 1) fundamental biology, 2) in situ analytics, 3) high performance computing, 4) automated science, 5) self-driving labs, 6) remote data management, 7) integrated mission biomonitoring, and 8) a Precision Space Health system. AI/ML approaches that can be integrated to support these domains include active learning, explainable AI, adaptive learning, causal inference, knowledge graphs, federated learning, transfer learning, and large language models. Finally, we present results from several current ML projects that are underway in the space biology field to address key challenges of small sample n, high feature count, heterogeneity, and sparse data. These include 1) connecting omics to phenotypic data using an ensemble model to infer causality of rodent liver health disruption, 2) usage of explainable ML to interrogate muscular underpinnings of muscle atrophy, 3) ML models analyzing and determining directed acyclic graphs of human health risk leveraging rodent bone datasets, 4) usage of large pre-trained models connecting biomedical knowledgebases with small spaceflight datasets to understand gene-to-gene interactions, and 5) a suite of benchmarked open science datasets enabling programmers to identify best algorithms to answer space biology questions.

space biology

Biological Research and Space Health Enabled by Machine Learning to Support Deep Space Missions

A key science goal of the NASA “Moon to Mars” campaign is to understand how biology responds to the Lunar, Martian, and deep space environments in order to advance fundamental knowledge, reduce risk, and support safe, productive human space missions. Through the powerful emerging computer science approaches of artificial intelligence (AI) and machine learning (ML), a paradigm shift has begun in biomedical science and engineered astronaut health systems, to enable Earth-independence and autonomy of mission operations. We present a decadal view of AI/ML architecture to support deep space mission goals, developed in concert with leaders in the field. We describe current AI/ML methods to support 1) fundamental biology, 2) in situ analytics, 3) high performance computing hardware, 4) automated science, 5) self-driving labs, 6) remote data management, 7) integrated real-time mission biomonitoring, and 8) a Precision Space Health system. Cutting-edge AI/ML approaches that can be integrated to support these domains include active learning, explainable AI, adaptive learning, causal inference, knowledge graphs, federated learning, transfer learning, and large language models. Finally, we present results from several current ML projects that are underway in the field to address key challenges of small sample n, high feature count, heterogeneity, and sparse data. These include 1) connecting omics data to phenotypic data using an ensemble model to infer causality of spaceflight rodent liver health disruption, 2) usage of explainable ML to interrogate the muscular underpinnings of spaceflight muscle atrophy, 3) ML models analyzing and determining directed acyclic graphs of human space health risk leveraging rodent bone datasets, 4) usage of large pre-trained models connecting biomedical knowledgebases with small spaceflight datasets to understand gene-to-gene interaction networks, and 5) a suite of benchmarked open science datasets (spaceflight mouse liver; radiation DNA damage) enabling programmers to identify the best ML algorithms to answer space biological science questions.

space biology

The Atacama Cosmology Telescope: Map-Based Noise Simulations for DR6

The increasing statistical power of cosmic microwave background (CMB) datasets requires a commensurate effort in understanding their noise properties. The noise in maps from ground-based instruments is dominated by large-scale correlations, which poses a modeling challenge. This paper develops novel models of the complex noise covariance structure in the Atacama Cosmology Telescope Data Release 6 (ACT DR6) maps. We first enumerate the noise properties that arise from the combination of the atmosphere and the ACT scan strategy. We then prescribe a class of Gaussian, map-based noise models, including a new wavelet-based approach that uses directional wavelet kernels for modeling correlated instrumental noise. The models are empirical, whose only inputs are a small number of independent realizations of the same region of sky. We evaluate the performance of these models against the ACT DR6 data by drawing ensembles of noise realizations. Applying these simulations to the ACT DR6 power spectrum pipeline reveals a ≥ 20% excess in the covariance matrix diagonal when compared to an analytic expression that assumes noise properties are uniquely described by their power spectrum. Along with our public code, mnms, this work establishes a necessary element in the science pipelines of both ACT DR6 and future ground-based CMB experiments such as the Simons Observatory (SO).

CMBR experiments

On the Prediction of Surface Melt Over the Greenland Ice Sheet at NWP and S2S Timescales

Greenland surface melt “events” – encompassing large swaths of the ice sheet over approximately 4-10 day periods – have occurred with increasing frequency in the 21st Century. Surface processes are significant to the ice sheet’s overall contribution to sea level. In addition, the sudden onset of widespread surface melt produces hazardous conditions for researchers and for communities adjacent to the ice sheet. The mechanisms producing melt events, their associated large-scale forcing, and timescales of interest are highly variable and poorly understood. Here we examine the predictability of melt events in a numerical weather prediction system (NWP) and a subseasonal-to-seasonal system (S2S). The NASA Global Modeling and Assimilation Office forward processing (GMAO-FP) is an NWP system that incorporates an evolving, global hybrid 4D ensemble–variational (4DEnVar) data assimilation system and a state-of-the-art model that is run twice-daily with forecasts out to five and ten days. The GMAO S2S system is a hindcast-anomaly, coupled-model prediction system initialized with an atmospheric analysis, and in situ and satellite ocean data assimilation. Both systems represent ice sheet surface hydrology and snowpack processes. We examine the performance of the systems over the period 2018-2023, a span that includes approximately nine significant events, against analyses and satellite-derived melt extent. In addition to extent, we examine related conditions and their role in the melt forecast, including the onset of both omega and Rex atmospheric blocking patterns, sea surface temperatures, sea ice extent, water vapor transport, precipitation, and surface preconditioning. We further examine the tendency for false-alarm events in each system and assess reasons for the forecast performance. North Atlantic atmospheric blocking patterns leading to surface melt are well forecast by the NWP system out to around eight days, while the S2S system generally underestimates the magnitude of such events at monthly lead times. These systems suggest that while short term forecasts of ice sheet melt and melt events are relatively robust, S2S predictions of surface conditions are highly dependent on the ability of a coupled system to generate the short-term atmospheric conditions associated with surface melt events.

Richard Cullather

Unbinned extraction of $γ$ from $B\to DK$ with normalizing flows

We introduce an unbinned method for extracting the CKM angle $γ$ from the decay chain $B^\pm \to (D \to K_S π^+ π^-) K^\pm$ using normalizing flows (NFs). The NFs, trained on $D$ decay data, learn a faithful continuous representation of the amplitude and strong phase variation over the $D\to K_Sπ^+π^-$ Dalitz plot whose fidelity improves with increased data sample sizes. With this input, the $B$ decay data can be used to extract the parameters $r_B$, $δ_B$, and $γ$. We test the method on Monte Carlo generated data, where it successfully recovers the injected value of $γ$ within uncertainties. The present implementation propagates statistical uncertainties from finite training data via an ensemble of independently trained flows, and does not attempt to capture the effects of systematic experimental errors. We explore two versions of the method that differ in how the trigonometric constraint on phase variation is encoded, and comment on the possible extension to Bayesian NFs, which would provide direct uncertainty estimates on the learned densities without requiring ensemble training.

Grossman, Yuval [Cornell U., LEPP]

The Impact of Model and Rainfall Forcing Errors on Characterizing Soil Moisture Uncertainty in Land Surface Modeling

The contribution of rainfall forcing errors relative to model (structural and parameter) uncertainty in the prediction of soil moisture is investigated by integrating the NASA Catchment Land Surface Model (CLSM), forced with hydro-meteorological data, in the Oklahoma region. Rainfall-forcing uncertainty is introduced using a stochastic error model that generates ensemble rainfall fields from satellite rainfall products. The ensemble satellite rain fields are propagated through CLSM to produce soil moisture ensembles. Errors in CLSM are modeled with two different approaches: either by perturbing model parameters (representing model parameter uncertainty) or by adding randomly generated noise (representing model structure and parameter uncertainty) to the model prognostic variables. Our findings highlight that the method currently used in the NASA GEOS-5 Land Data Assimilation System to perturb CLSM variables poorly describes the uncertainty in the predicted soil moisture, even when combined with rainfall model perturbations. On the other hand, by adding model parameter perturbations to rainfall forcing perturbations, a better characterization of uncertainty in soil moisture simulations is observed. Specifically, an analysis of the rank histograms shows that the most consistent ensemble of soil moisture is obtained by combining rainfall and model parameter perturbations. When rainfall forcing and model prognostic perturbations are added, the rank histogram shows a U-shape at the domain average scale, which corresponds to a lack of variability in the forecast ensemble. The more accurate estimation of the soil moisture prediction uncertainty obtained by combining rainfall and parameter perturbations is encouraging for the application of this approach in ensemble data assimilation systems.

Maggioni, V.

Evaluating the Global Water and Energy Cycles in Reanalyses Using Ensemble Spread

Reanalyses provide continuous observation based global data over weather and climate time scales through observational assimilation. Ensemble forecasts are a component of modern data assimilation systems, providing information to the analysis process. Storing this spread over the duration of a climate reanalysis will provide insight to the variations of the model that underlies the reanalysis data. Global water and energy cycles (WECs) are a critical aspect of both weather and climate, yet there are some large gaps in the observation of key components as well as their representation in models and reanalyses. NASA’s Water and Energy cycle Studies (NEWS) program is developing a corrected and balanced budget data set that is closed at monthly and regional scales during the 21st century, so far. In this paper, we will evaluate reanalysis WEC components against the initial NEWS collection of observation, but integrate the analysis spread from the reanalysis data assimilation as a new perspective on the evaluation. The initial evaluation will cover a few more recent years using ERA5 10 member spread. As the GMAO’s production of the Modern-Era Retrospective analysis for Research and Applications 21st Century (MERRA-21C) progresses, the effort will include the MERRA-21C 32 member spread. Ultimately, we will explore the potential of the ensemble spread as a measure of uncertainty in the reanalysis.

Michael Bosilovich

Influence of Leaf Area Index Prescriptions on Simulations of Heat, Moisture, and Carbon Fluxes

Leaf-area index (LAI), the total one-sided surface area of leaf per ground surface area, is a key component of land surface models. We investigate the influence of differing, plausible LAI prescriptions on heat, moisture, and carbon fluxes simulated by the Community Atmosphere Biosphere Land Exchange (CABLEv1.4b) model over the Australian continent. A 15-member ensemble monthly LAI data-set is generated using the MODIS LAI product and gridded observations of temperature and precipitation. Offline simulations lasting 29 years (1980-2008) are carried out at 25 km resolution with the composite monthly means from the MODIS LAI product (control simulation) and compared with simulations using each of the 15-member ensemble monthly-varying LAI data-sets generated. The imposed changes in LAI did not strongly influence the sensible and latent fluxes but the carbon fluxes were more strongly affected. Croplands showed the largest sensitivity in gross primary production with differences ranging from -90 to 60 %. PFTs with high absolute LAI and low inter-annual variability, such as evergreen broadleaf trees, showed the least response to the different LAI prescriptions, whilst those with lower absolute LAI and higher inter-annual variability, such as croplands, were more sensitive. We show that reliance on a single LAI prescription may not accurately reflect the uncertainty in the simulation of the terrestrial carbon fluxes, especially for PFTs with high inter-annual variability. Our study highlights that the accurate representation of LAI in land surface models is key to the simulation of the terrestrial carbon cycle. Hence this will become critical in quantifying the uncertainty in future changes in primary production.

land-surface modeling

Downscaled Daily 1 km Climate Data (NEX-GDDP-CMIP6) for Southeast Texas. Full ensemble of downscaled CMIP6 climate projections at 1 km daily resolution.

For the SETx-UIFL, the daily NASA Earth Exchange Global Daily Downscaled Projections (NEX-GDDP-CMIP6) dataset climate projections were downscaled from approximately 27 km to 1 km. The SETx dataset provides very high-resolution climate data for the historical period (1950–2014) and future scenarios derived from CMIP6 global models under the four Tier 1 Shared Socioeconomic Pathways (SSPs 1.26, 2.45, 3.70, and 5.85), developed for the IPCC Sixth Assessment Report. A subset of ten NEX-GDDP-CMIP6 models was selected to represent a balance of model families, climate sensitivities, and availability across scenarios, ensuring a diverse and reliable ensemble for regional analysis. Selected models: BCC-CSM2-MR, CESM2, CMCC-ESM2, CNRM-ESM2-1, EC-Earth3, FGOALS-g3, GFDL-CM4, MPI-ESM1-2-HR, MRI-ESM2-0, NorESM2-MM. Daily variables downscaled include tasmax, tasmin, tas, pr, hurs, huss, rsds, rlds, and sfcWind.

Persad, Geeta

Lessons Learned from Assimilating Altimeter Data into a Coupled General Circulation Model with the GMAO Augmented Ensemble Kalman Filter

Satellite altimetry measurements have provided global, evenly distributed observations of the ocean surface since 1993. However, the difficulties introduced by the presence of model biases and the requirement that data assimilation systems extrapolate the sea surface height (SSH) information to the subsurface in order to estimate the temperature, salinity and currents make it difficult to optimally exploit these measurements. This talk investigates the potential of the altimetry data assimilation once the biases are accounted for with an ad hoc bias estimation scheme. Either steady-state or state-dependent multivariate background-error covariances from an ensemble of model integrations are used to address the problem of extrapolating the information to the sub-surface. The GMAO ocean data assimilation system applied to an ensemble of coupled model instances using the GEOS-5 AGCM coupled to MOM4 is used in the investigation. To model the background error covariances, the system relies on a hybrid ensemble approach in which a small number of dynamically evolved model trajectories is augmented on the one hand with past instances of the state vector along each trajectory and, on the other, with a steady state ensemble of error estimates from a time series of short-term model forecasts. A state-dependent adaptive error-covariance localization and inflation algorithm controls how the SSH information is extrapolated to the sub-surface. A two-step predictor corrector approach is used to assimilate future information. Independent (not-assimilated) temperature and salinity observations from Argo floats are used to validate the assimilation. A two-step projection method in which the system first calculates a SSH increment and then projects this increment vertically onto the temperature, salt and current fields is found to be most effective in reconstructing the sub-surface information. The performance of the system in reconstructing the sub-surface fields is particularly impressive for temperature, but not as satisfactory for salt.

Keppenne, Christian

Dataset for manuscript "Equipartition and the temperature of maximum density of TIP4P/2005 water"

We simulate TIP4P/2005 water in the temperature range of 257 K to 318 K with time-steps 0.25, 0.50, 1.00, 2.00, and 4.00 fs. The density-temperature behavior obtained using 0.25 or 0.50 fs are in excellent agreement with each other but differ from those obtained using time-steps that have been shown earlier to lead to a breakdown of equipartition. The temperature of maximum density (TMD) is 277.15 K with time-step 0.25 or 0.50 fs, but is shifted to progressively lower values for longer time-steps, a trend that holds for different thermostat/barostat combinations. Enhancing the water-water dispersion interaction, as has been recommended for simulating disordered proteins in TIP4P/2005, degrades the description of the liquid-vapor phase envelope. We present a simple physically transparent reasoning to highlight the separation of the time-scales between translational and rotational motion. We also develop a metric, Chi, that we term the equipartition anomaly, to detect equipartition violations in simulations that include molecules that are treated as rigid objects. Calculating Chi is shown to be straightforward and sensitive to equipartition violations. A key takeaway from this study is that using sufficiently short time-steps (less than or equal to 0.5 fs) to preserve equipartition is essential for obtaining meaningful liquid water properties and for producing reliable simulation data, as correct-ensemble sampling is fundamental to ensure reproducibility across codes and simulation alogrithms. The included dataset provides the raw data used in the preparation of the graphs noted in the manuscript.

36 MATERIALS SCIENCE

Analysis of instrumentation error effects on the identification accuracy of aircraft parameters

An analytical investigation is presented of the effect of unmodeled measurement system errors on the accuracy of aircraft stability and control derivatives identified from flight test data. Such error sources include biases, scale factor errors, instrument position errors, misalignments, and instrument dynamics. Two techniques (ensemble analysis and simulated data analysis) are formulated to determine the quantitative variations to the identified parameters resulting from the unmodeled instrumentation errors. The parameter accuracy that would result from flight tests of the F-4C aircraft with typical quality instrumentation is determined using these techniques. It is shown that unmodeled instrument errors can greatly increase the uncertainty in the value of the identified parameters. General recommendations are made of procedures to be followed to insure that the measurement system associated with identifying stability and control derivatives from flight test provides sufficient accuracy.

Sorensen, J. A.

Evaluating the Utility of Remotely-Sensed Soil Moisture Retrievals for Operational Agricultural Drought Monitoring

Soil moisture is a fundamental data source used by the United States Department of Agriculture (USDA) International Production Assessment Division (IPAD) to monitor crop growth stage and condition and subsequently, globally forecast agricultural yields. Currently, the USDA IPAD estimates surface and root-zone soil moisture using a two-layer modified Palmer soil moisture model forced by global precipitation and temperature measurements. However, this approach suffers from well-known errors arising from uncertainty in model forcing data and highly simplified model physics. Here we attempt to correct for these errors by designing and applying an Ensemble Kalman filter (EnKF) data assimilation system to integrate surface soil moisture retrievals from the NASA Advanced Microwave Scanning Radiometer (AMSR-E) into the USDA modified Palmer soil moisture model. An assessment of soil moisture analysis products produced from this assimilation has been completed for a five-year (2002 to 2007) period over the North American continent between 23degN - 50degN and 128degW - 65degW. In particular, a data denial experimental approach is utilized to isolate the added utility of integrating remotely-sensed soil moisture by comparing EnKF soil moisture results obtained using (relatively) low-quality precipitation products obtained from real-time satellite imagery to baseline Palmer model runs forced with higher quality rainfall. An analysis of root-zone anomalies for each model simulation suggests that the assimilation of AMSR-E surface soil moisture retrievals can add significant value to USDA root-zone predictions derived from real-time satellite precipitation products.

Bolten, John D.

Exploring Water System Vulnerabilities in California's Central Valley Under the Late Renaissance Megadrought and Climate Change

Abstract California faces cycles of drought and flooding that are projected to intensify, but these extremes may impact water users across the state differently due to the region's natural hydroclimate variability and complex institutional framework governing water deliveries. To assess these risks, this study introduces a novel exploratory modeling framework informed by paleo and climate‐change based scenarios to better understand how impacts propagate through the Central Valley's complex water system. A stochastic weather generator, conditioned on tree‐ring data, produces a large ensemble of daily weather sequences conditioned on drought and flood conditions under the Late Renaissance Megadrought period (1550–1580 CE). Regional climate changes are applied to this weather data and drive hydrologic projections for the Sacramento, San Joaquin, and Tulare Basins. The resulting streamflow ensembles are used in an exploratory stress test using the California Food‐Energy‐Water System model, a highly resolved, daily model of water storage and conveyance throughout California's Central Valley. Results show that megadrought conditions lead to unprecedented reductions in inflows and storage at major California reservoirs. Both junior and senior water rights holders experience multi‐year periods of curtailed water deliveries and complete drawdowns of groundwater assets. When megadrought dynamics are combined with climate change, risks for unprecedented depletion of reservoir storage and sustained curtailment of water deliveries across multiple years increase. Asymmetries in risk emerge depending on water source, rights, and access to groundwater banks.

Gupta, Rohini S. [School of Civil and Environmenta

An interplanetary magnetic field ensemble at 1 AU

A method for calculation ensemble averages from magnetic field data is described. A data set comprising approximately 16 months of nearly continuous ISEE-3 magnetic field data is used in this study. Individual subintervals of this data, ranging from 15 hours to 15.6 days comprise the ensemble. The sole condition for including each subinterval in the averages is the degree to which it represents a weakly time-stationary process. Averages obtained by this method are appropriate for a turbulence description of the interplanetary medium. The ensemble average correlation length obtained from all subintervals is found to be 4.9 x 10 to the 11th cm. The average value of the variances of the magnetic field components are in the approximate ratio 8:9:10, where the third component is the local mean field direction. The correlation lengths and variances are found to have a systematic variation with subinterval duration, reflecting the important role of low-frequency fluctuations in the interplanetary medium.

Matthaeus, W. H.