Search NASA⌕ Search

SEARCH · Search NASA

Results for “spatial data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 253 records · Page 14

Spatial Proteomics towards cellular Resolution

Introduction: Spatial biology is an emerging interdisciplinary field facilitating biological discoveries through the use of spatial omics technologies. Recent advancements in spatial transcriptomics, spatial genomics (e.g. genetic mutations and epigenetic marks), multiplexed immunofluorescence, and spatial metabolomics/lipidomics have enabled high-resolution spatial profiling of gene expression, genetic variation, protein expression, and metabolites/lipids profiles in tissue. These developments contribute to a deeper understanding of the spatial organization within tissue microenvironments at the molecular level. Areas covered: This report provides an overview of the untargeted, bottom-up mass spectrometry (MS)-based spatial proteomics workflow. It highlights recent progress in tissue dissection, sample processing, bioinformatics, and liquid chromatography (LC)-MS technologies that are advancing spatial proteomics toward cellular resolution. Expert opinion: The field of untargeted MS-based spatial proteomics is rapidly evolving and holds great promise. To fully realize the potential of spatial proteomics, it is critical to advance data analysis and develop automated and intelligent tissue dissection at the cellular or subcellular level, along with high-throughput LC-MS analyses of thousands of samples. In conclusion, achieving these goals will necessitate significant advancements in tissue dissection technologies, LC-MS instrumentation, and computational tools.

59 BASIC BIOLOGICAL SCIENCES↗

Spectrum and extension of the inverse-Compton emission of the Crab Nebula from a combined Fermi -LAT and H.E.S.S. analysis

The Crab Nebula is a unique laboratory for studying the acceleration of electrons and positrons through their non-thermal radiation. Observations of very-high-energy γ rays from the Crab Nebula have provided important constraints for modelling its broadband emission. We present the first fully self-consistent analysis of the Crab Nebula’s γ-ray emission between 1 GeV and ∼100 TeV, that is, over five orders of magnitude in energy. Using the open-source software package GAMMAPY, we combined 11.4 yr of data from the Fermi Large Area Telescope and 80 h of High Energy Stereoscopic System (H.E.S.S.) data at the event level and provide a measurement of the spatial extension of the nebula and its energy spectrum. We find evidence for a shrinking of the nebula with increasing γ-ray energy. Furthermore, we fitted several phenomenological models to the measured data, finding that none of them can fully describe the spatial extension and the spectral energy distribution at the same time. Especially the extension measured at TeV energies appears too large when compared to the X-ray emission. Our measurements probe the structure of the magnetic field between the pulsar wind termination shock and the dust torus, and we conclude that the magnetic field strength decreases with increasing distance from the pulsar. We complement our study with a careful assessment of systematic uncertainties.

79 ASTRONOMY AND ASTROPHYSICS↗

Quantitative phenotyping of crop roots with spectral electrical impedance tomography: a rhizotron study with optimized measurement design

Background: Root systems are key contributors to plant health, resilience, and, ultimately, yield of agricultural crops. To optimize plant performance, phenotyping trials are conducted to breed plants with diverse root traits. However, traditional analysis methods are often labour-intensive and invasive to the root system, therefore limiting high-throughput phenotyping. Spectral electrical impedance tomography (sEIT) could help as a non-invasive and cost-efficient alternative to optical root analysis, potentially providing 2D or 3D spatio-temporal information on root development and activity. Although impedance measurements have been shown to be sensitive to root biomass, nutrient status, and diurnal activity, only few attempts have been made to employ tomographic algorithms to recover spatially resolved information on root systems. In this study, we aim to establish relationships between tomographic electrical polarization signatures and root traits of different fine root systems (maize, pinto bean, black bean, and soy bean) under hydroponic conditions. Results: Our results show that, with the use of an optimized data acquisition scheme, sEIT is capable of providing spatially resolved information on root biomass and root surface area for all investigated root systems. We found strong correlations between the total polarization strength and the root biomass (R 2 = 0.82) and root surface area (R 2 = 0.8). Our findings suggest that the captured polarization signature is dominated by cell-scale polarization processes. Additionally, we demonstrate that the resolution characteristics of the measurement scheme can have a significant impact on the tomographic reconstruction of root traits. Conclusion: Our findings showcase that sEIT is a promising tool for the tomographic reconstruction of root traits in high-throughput root phenotyping trials and should be evaluated as a substitute for traditional, often time-consuming, root characterization methods.

59 BASIC BIOLOGICAL SCIENCES↗

UFNet: Joint U-Net and Fully Connected Neural Network to Bias Correct Precipitation Predictions from

Paper information. Shuang Yu, Indrasis Chakraborty, Gemma J. Anderson, Donald D. Lucas, Yannic Lops, and Daniel Galea. UFNet: Joint U-Net and fully connected neural network to bias correct precipitation predictions from climate models. Artificial Intelligence for the Earth Systems, 2024. Overview. This work develops the UFNet methodology to correct E3SM historical precipitation projection bias. The UFNet deep learning framework consists of a two-part architecture: a U-Net convolutional network to capture the spatiotemporal distribution of precipitation and a fully connected network to capture the distribution of higher-order statistics. The joint network, termed UFNet, can simultaneously improve the spatial structure of the modeled precipitation and capture the distribution of extreme precipitation values. Below we provide guidance for applying UFNet to correct the Energy Exascale Earth System Model (E3SM; Golaz et al. 2019) daily precipitation projection over the contiguous United States (CONUS). Getting started 1. Obtain the historical climate simulation and observation data. The E3SM historical simulation data are available through https://aims2.llnl.gov/search/cmip6/. The CPC unified gauge-based analysis of daily precipitation can be found through https://psl.noaa.gov/data/gridded/data.cpc.globalprecip.html. The ECMWF atmospheric reanalysis of the 20th century (ERA-20C) data are available through https://www.ecmwf.int/en/forecasts/datasets/reanalysis-datasets/era-20c. The spatial resolution of E3SM and observed datasets are both regridded to a common 1° resolution grid using conservative interpolation. The regridded E3SM, CPC and ERA-20C with 1° resolution can be found throught ./data/. 2. Train the fully connected network (DNN) Python train_dnn.py 3. Train the UFNet Python train_ufnet.py 4. Evaluation and compared with the baseline Python evaluation.py

Lucas, DonaldD↗

A Particle-in-Cell Method for Plasmas with a Generalized Momentum Formulation, Part II: Enforcing the Lorenz Gauge Condition

In a previous paper Christlieb et al. (A particle-in-cell method for plasmas with a generalized momentum formulation, part I: Model formulation, 2024), we developed a new particle-in-cell (PIC) method for the relativistic Vlasov–Maxwell system in which the electromagnetic fields and the equations of motion for the particles were cast in terms of scalar and vector potentials through a Hamiltonian formulation. This new method evolved the potentials under the Lorenz gauge using integral equation methods. New methods to construct spatial derivatives of the potentials that converge at the same rates as the fields were also presented. The new particle method was compared against standard explicit discretizations, including the well-known FDTD-PIC method, for a range of applications involving sheaths and particle beams. Here, this paper extends this new class of methods by focusing on the enforcement the Lorenz gauge condition in both exact and approximate forms using co-located meshes. A time-consistency property of the proposed field solver for the vector potential form of Maxwell’s equations is established, which is shown to preserve the equivalence between the semi-discrete Lorenz gauge condition and the analogous semi-discrete continuity equation. Using this property, we present three methods to enforce a semi-discrete gauge condition. The first method introduces an update for the continuity equation that is consistent with the discretization of the Lorenz gauge condition. Both the finite difference and spectral implementations satisfy this discrete gauge condition to machine precision. The second approach we propose enforces a semi-discrete continuity equation using the boundary integral solution to the field equations. The potential benefit of this approach is that it eliminates spatial derivatives that appear on the particle data, namely the current density, which is often calculated by linear combinations of low-order spline basis functions. This method is ideally suited to boundary integral equation methods that invert multi-dimensional operators without dimensional splitting techniques and will be the subject of future work. The third approach introduces a gauge correcting method that makes direct use of the gauge condition to modify the scalar potential and uses local maps for both the charge and current densities. This results in a gauge error, as the maps do not enforce the continuity equation. The vector potential coming from the current density is taken to be exact, and using the Lorenz gauge, we compute a correction to the scalar potential that makes the two potentials satisfy the gauge condition. This method also enforces the gauge condition to machine precision. We demonstrate two of the proposed methods in the context of periodic domains. Problems defined on bounded domains, including those with complex geometric features remain an ongoing effort. However, this work shows that it is possible to design computationally efficient methods that can effectively enforce the Lorenz gauge condition in a non-staggered PIC formulation.

97 MATHEMATICS AND COMPUTING↗

High-throughput micro-scale bandgap mapping for perovskite-inspired materials with complex composition space

Abstract To realize the full promise of high-throughput experimental workflows, the rate of sample synthesis must be matched by that of characterization. Of growing interest are contactless optical techniques that can rapidly measure material homogeneity and properties. Here, we present a hyperspectral imaging method to measure local optical bandgap distributions within samples, utilizing spatially-resolved reflectance spectra coupled with automated data analysis. We collect approximately one million optical bandgap data across the compositional space of Cs 3 (Bi x Sb 1-x ) 2 (Br y I 1-y ) 9 perovskite-inspired materials. Our results show non-monotonic bandgap variations (i.e., bandgap bowing) along six composition gradient sequences, in addition to identifying samples with multiple bandgaps in statistics. High-throughput transient absorption spectroscopy reveals that within these compositions, the depletion of the ground state carriers to excited states occurred at discrete energy levels with independent carrier dynamics, consistent with the bandgap observation and indicative of phase separation. This work demonstrates the potential for rapid optical measurements to assess material quality and homogeneity in a high-throughput experimental setting, supporting screening and recipe optimization of optoelectronic material candidates with desired carrier dynamics and optical properties.

Science & Technology - Other Topics↗

Efficient and generalizable nested Fourier-DeepONet for three-dimensional geological carbon sequestration

Geological carbon sequestration (GCS) involves injecting CO2 into subsurface geological formationsfor permanent storage. Numerical simulations could guide decisions in GCS projects by predictingCO 2 migration pathways and the pressure distribution in storage formation. However, these simula-tions are often computationally expensive due to highly coupled physics and large spatial-temporalsimulation domains. Surrogate modelling with data-driven machine learning has become a promis-ing alternative to accelerate physics-based simulations. Among these, the Fourier neural operator(FNO) has been applied to three-dimensional synthetic subsurface models. Despite its good accuracyin simulating CO 2 plume migration, it requires large computational resources in training and alsolacks generalizability. Here, to further improve performance, we have developed a nested Fourier-DeepONet by combining the expressiveness of the FNO with the modularity of a deep operatornetwork (DeepONet). This new framework is twice as efficient as a nested FNO for training and has atleast 80% lower GPU memory requirement due to its flexibility to treat temporal coordinates sepa-rately. These performance improvements are achieved without compromising prediction accuracy.In addition, the generalization and extrapolation ability of nested Fourier-DeepONet beyond thetraining range has been thoroughly evaluated. Nested Fourier-DeepONet outperformed the nestedFNO for extrapolation in time with more than 50% reduced error. It also exhibited good extrapolationaccuracy beyond the training range in terms of reservoir properties, number of wells, and injectionrate.

Lee, Jonathan E. [Department of Chemical and Envir↗

Scalable edge clustering of dynamic graphs via weighted line graphs

Timestamped relational datasets consisting of records (or connections) between pairs of entities are ubiquitous in network science. For applications like peer-to-peer communication, email, various social network interactions, and computer network security, it is useful to organize these records into groups based on how and when they are occurring. Weighted line graphs offer a natural way to model how records are related in such datasets but for large real-world graph topologies, building and utilizing the line graph is prohibitively expensive. Here, we present the framework to cluster the edges of a dynamic graph via the associated line graph that contains two major contributions. The first is a method to work with the line graph implicitly and the second is a distributed scale implementation of an agglomerative hierarchical graph clustering algorithm. We outline a novel hierarchical dynamic graph edge clustering approach that efficiently breaks massive relational datasets into small sets of edges containing events at various timescales. This is in stark contrast to traditional graph clustering algorithms that prioritize highly connected (clique-like) community structures. Our approach relies on constructing a sufficient subgraph of a weighted line graph and applying a hierarchical agglomerative clustering. This approach is related to scalable techniques from spatial clustering, nonlinear-dimension reduction, topological data analysis, and draws particular inspiration from HDBSCAN. As an edge clustering, this method yields an overlapping node clustering. Our algorithm is parallelizable and we demonstrate efficient clustering of a billion-scale, real-world dynamic graph into small edge sets that correlate in topology and time. The entire clustering process for a graph with tens of billions of edges takes just a few minutes of run time on 256 nodes of a distributed compute environment. We argue how the output of the edge clustering is useful for a multitude of data visualization and powerful machine learning tasks, both involving the original massive dynamic graph data and metadata associated with the nodes and edges. Finally, we describe how this approach can be extended to dynamic hypergraphs and dynamic graphs/hypergraphs with unstructured data living on vertices and edges.

Data Analysis↗

Searching for synchrotron emission from the geminga TeV halo using the planck satellite

Pulsars convert a significant fraction of their total spin-down power into very high-energy electrons, leading to the formation of TeV halos. While these halos are well characterized at TeV energies, it remains unclear whether pulsars also accelerate electrons efficiently at lower energies and how these particles propagate through their surrounding environments. We aim to test whether synchrotron emission from $\sim 50$–$300 \, \textrm{GeV}$ electrons around the Geminga pulsar can be detected in the frequency range observed by the Planck satellite. This would help constrain low-energy particle acceleration and diffusion in the vicinity of pulsars. We model the expected synchrotron emission from Geminga’s TeV halo based on various diffusion and injection spectrum scenarios and compare these predictions to publicly available multi-frequency Planck data. We find no conclusive evidence of spatially extended synchrotron emission associated with Geminga in any of Planck’s frequency bands. Our calculations show that even under favorable diffusion and injection conditions, the predicted synchrotron flux lies well below Planck’s measured background levels.

79 ASTRONOMY AND ASTROPHYSICS↗

Leveraging Optimal Sparse Sensor Placement to Aggregate a Network of Digital Twins for Nuclear Subsystems

Nuclear power plants (NPPs) require continuous monitoring of various systems, structures, and components to ensure safe and efficient operations. The critical safety testing of new fuel compositions and the analysis of the effects of power transients on core temperatures can be achieved through modeling and simulations. They capture the dynamics of the physical phenomenon associated with failure modes and facilitate the creation of digital twins (DTs). Accurate reconstruction of fields of interest (e.g., temperature, pressure, velocity) from sensor measurements is crucial to establish a two-way communication between physical experiments and models. Sensor placement is highly constrained in most nuclear subsystems due to challenging operating conditions and inherent spatial limitations. This study develops optimized data-driven sensor placements for full-field reconstruction within reactor and steam generator subsystems of NPPs. Optimized constrained sensors reconstruct field of interest within a tri-structural isotropic (TRISO) fuel irradiation experiment, a lumped parameter model of a nuclear fuel test rod and a steam generator. The optimization procedure leverages reduced-order models of flow physics to provide a highly accurate full-field reconstruction of responses of interest, noise-induced uncertainty quantification and physically feasible sensor locations. Accurate sensor-based reconstructions establish a foundation for the digital twinning of subsystems, culminating in a comprehensive DT aggregate of an NPP.

42 ENGINEERING↗

The DECam MAGIC Survey: A Wide-field Photometric Metallicity Study of the Sculptor Dwarf Spheroidal Galaxy

The metallicity distribution function (MDF) and internal chemical variations of a galaxy are fundamental to understand its formation and assembly history. In this work, we analyze photometric metallicities for 3883 stars over 7 half-light radii (rh) in the Sculptor (Scl) dwarf spheroidal (dSph) galaxy, using new narrowband imaging data from the Mapping the Ancient Galaxy in CaHK (MAGIC) survey conducted with the Dark Energy Camera (DECam) at the 4 m Blanco Telescope. This work demonstrates the scientific potential of MAGIC using the Scl dSph galaxy, one of the most well-studied satellites of the Milky Way. Our sample ranges from [Fe/H] ≈ –4.0 to [Fe/H] ≈ –0.6, includes six new extremely metal-poor candidates ([Fe/H] ≤ –3.0), and is almost 3 times larger than the largest spectroscopic metallicity data set in the Scl dSph. Our spatially unbiased sample of metallicities provides a more accurate representation of the MDF, revealing a more metal-rich peak than observed in the most recent spectroscopic sample. It also reveals a break in the metallicity gradient, with a strong change in the slope: from −3.26 ± 0.18 dex deg −1 for stars inside ∼1 rh to −0.55 ± 0.26 dex deg −1 for the outer part of the Scl dSph. Our study demonstrates that combining photometric metallicity analysis with the wide field of view of DECam offers an efficient and unbiased approach for studying the stellar populations of dwarf galaxies in the Local Group.

79 ASTRONOMY AND ASTROPHYSICS↗

Probing the γ -Ray Emission Origin of Two Star-forming Galaxies NGC 2403 and NGC 3424 with the Fermi-LAT

Star-forming galaxies (SFGs) are a subclass of γ-ray emitters, and a correlation between their γ-ray luminosity (L γ ) and the total infrared (IR) luminosity (L IR ) has been established based on the Fermi Large Area Telescope (LAT) data. NGC 2403 and NGC 3424 have been reported as outliers in the L γ –L IR correlation with light curves showing significant variability, which contrasts with the temporally stable γ-ray emission in other SFGs, originating primarily from cosmic rays interacting with interstellar medium. In this study, we reanalyze the γ-ray emission in the directions of NGC 2403 and NGC 3424 using more than 16.5 yr Fermi-LAT data. NGC 3424 is found to be spatially coincident with the detected γ-ray source, while NGC 2403 is significantly offset from the nearest γ-ray source, suggesting an implausible association. We confirm the previously reported variability of both γ-ray sources and the significant deviation from the L γ –L IR correlation when assuming an association of both γ-ray sources with the two galaxies. Our findings lend further support to the interpretation that their γ-ray emission is driven primarily by alternative radiative processes—rather than by star formation activity—such as the ejecta of the Type IIP supernova SN 2004dj in NGC 2403 interacting with a surrounding high-density shell and an obscured active galactic nucleus in NGC 3424.

Liu, Linjie [Chinese Academy of Sciences (CAS), Ku↗

Analysis of carbon capture at cellulosic biorefineries

The large-scale production of cellulosic biofuels would involve spatially distributed systems including biomass fields, logistics networks and biorefineries. Better understanding of the interactions between landscape-related decisions and the design of biorefineries with carbon capture and storage (CCS) in a supply chain context is needed to enable efficient systems. Here we analyse the cost and greenhouse gas mitigation potential for cellulosic biofuel supply chains in the US Midwest using realistic spatially explicit land availability and crop productivity data and consider fuel conversion technologies with detailed CCS design for their associated CO2 streams.

carbon capture and storage (CCS)↗

Learning the boundary-to-domain mapping using Lifting Product Fourier Neural Operators for partial differential equations

Neural operators such as the Fourier Neural Operator (FNO) have been shown to provide resolution-independent deep learning models that can learn mappings between function spaces. For example, an initial condition can be mapped to the solution of a partial differential equation (PDE) at a future time-step using a neural operator. Despite the popularity of neural operators, their use to predict solution functions over a domain given only data over the boundary (such as a spatially varying Dirichlet boundary condition) remains unexplored. In this paper, we refer to such problems as boundary-to-domain problems; they have a wide range of applications in areas such as fluid mechanics, solid mechanics, heat transfer etc. We present a novel FNO-based architecture, named Lifting Product FNO (or LP-FNO) which can map arbitrary boundary functions defined on the lower-dimensional boundary to a solution in the entire domain. Specifically, two FNOs defined on the lower-dimensional boundary are lifted into the higher dimensional domain using our proposed lifting product layer. We demonstrate the efficacy and resolution independence of the proposed LP-FNO for the 2D Poisson equation.

Kashi, Aditya↗

Pixel-Registered Multimodal Synchrotron XRF and FTIR Microscopies Reveal Salinity Stress Response Mechanisms in Pistachio

Background: Salinity is a major abiotic stress that negatively affects nearly all plant species at all stages of growth. Drought and poor-quality irrigation cause high soil salinity and salt accumulation via evaporation, reducing crop productivity. Despite its critical importance, the spatial localization of salt ions and associated biochemical changes within plants experiencing high salinity remains largely unknown. In this study, we developed a multimodal imaging pipeline to understand the impact of salinity on the pistachio rootstock UCB-1 (Pistacia atlantica x Pistacia integerrima). We directly link biochemical fingerprints in stem tissue architecture with salt ion localization to provide insights into the strategies pistachio uses to tolerate salinity. Results: We observed that Pistacia spp. exposed to high salt conditions accumulated Ca, Si, Cl, Al and Mg as hotspots within the pith, compared to the control (of which only Ca and Al co-locate). In contrast, there was a decrease in K between the control and salinity treatment. Hotspots of amide I and II were present in the cortex and pith of the salinity treated sample. Additionally, the salinity treatment resulted in an increased abundance of pectin and carbohydrates within the pith compared to the control, and the abundance of esters/carboxylic acid was greater in the salinity treatment. Conclusions: We determined that Cl and K, S and P, and biochemical components polysaccharide and pectin, esters and carboxylic acid, amide I and cellulose are the strongest drivers of salinity- treatment induced variability. In the cortex and phloem/xylem, a negative K-Ca correlation decreases in the salinity treatment. Several hotspots of elements and amide I (proteins) appear under salinity treatment, particularly in the cortex, suggesting an increase in the production of stress-related proteins (in response to high Cl) and/or structural proteins (i.e. Ca). Together, these results indicate that pistachio responds to salinity through ion compartmentalization coupled with a targeted biochemical adjustment, rather than a broadscale tissue-wide response. Overall, these novel, spatially resolved pixel-registered multimodal imaging data provide an enabling platform to understand the mechanisms of salinity tolerance in Pistacia spp and can be broadly applied to studying stress-related phenotype response in various plant tissues.

FTIR spectromicroscopy↗

Integration of fixed-frequency and FM-CW (frequency-modulated continuous-wave) reflectometers for coincident turbulence measurements on LTX- β (Lithium Tokamak eXperiment- β )

The fixed-frequency and frequency-modulated continuous-wave (FM-CW) reflectometers on LTX-β (Lithium Tokamak eXperiment-β) have been configured to use the same transmission lines and antenna arrays for coincident views of the core and edge plasma. The fixed-frequency channels (13.1–20.5 and 20–40 GHz, tunable between discharges) provide time-resolved measurements of density fluctuations, while the FM-CW channels (13.1–20.2 and 19.5–33.5 GHz) measure the density profile and fluctuations, with high spatial resolution and a sampling rate determined by the frequency sweep interval (5 μs). Data from both reflectometers are synchronously acquired to simultaneously leverage the wide bandwidth and high spatial resolution of the respective systems. Experiments showed that mutual crosstalk interference is momentary and does not diminish the capability of either system. Spectral analysis indicated broad power spectra (several hundreds of kHz) and suggests that the signals from the FM-CW system are consistent with under-sampled fixed-frequency signals. Radial correlations were explored using data from the two reflectometers, as well as from the FM-CW system alone. The core channels showed high levels of agreement between these two comparisons, suggesting that the data from the reflectometers are interchangeable for statistical estimates. For the edge channels, comparisons using data from the FM-CW reflectometer alone showed significant decorrelation due to time lag caused by the finite frequency up-sweep duration. Alternatively, this effect is eliminated when cross-correlating data from the different reflectometers. Finally, these results highlight the advantages of operating the fixed-frequency and FM-CW reflectometers in this manner, where the combined system can overcome the limitations of each separate system.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Ground surface temperature derived Snow Cover Properties, Seward Peninsula, Alaska, 2019-2023

Snow-ground interface temperatures have been collected at the Teller mile marker 27 and Kougarok mile marker 64 field sites on the Seward Peninsula, Alaska from 2019 through 2023 (with data missing from Fall 2020 through Summer 2021 due to COVID). Temperatures were measured using iButton Link DS1921G-F5# Thermochron miniature temperature sensors and Tinytag TGP-4017 internal sensors deployed across the Kougarok 64 and Teller 27 field sites. These sensors are a cost-efficient way to collect snow-ground interface temperatures at a high spatial resolution, and when paired with air temperature data these measurements can provide insight into fine-scale variability in snowpack characteristics across the study sites. From this data, snow process metrics were calculated at each sensor location based on the methods outlined in Staub and Delaloye, 2017. Metrics are calculated daily for each sensor as well as over the entire season. These metrics include ground surface temperature (°C), the number of days under snow cover (number of days), the insulation effect of snow (unitless), the length of the transitional snow periods (number of days), as well as intermediaries such as temperature variability. Calculating these snow processes relies on the assumption that when snow covers a temperature sensor, it is buffered from diurnal fluctuations in air temperature by the insulating snow layer. More information on the calculated metrics can be found in the User Guide of this dataset, as well as in Staub and Delaloye’s 2017 publication Using Near-Surface Ground Temperature Data to Derive Snow Insulation and Melt Indices for Mountain Permafrost Applications. This dataset includes one daily and one seasonal *.csv file of metrics for every year of data, a daily and a seasonal *.csv data dictionary, and one User Guide document (*.pdf) describing data collection and processing.The Next-Generation Ecosystem Experiments: Arctic (NGEE Arctic), was a research effort to reduce uncertainty in Earth System Models by developing a predictive understanding of carbon-rich Arctic ecosystems and feedbacks to climate. NGEE Arctic was supported by the Department of Energy's Office of Biological and Environmental Research.The NGEE Arctic project had two field research sites: 1) located within the Arctic polygonal tundra coastal region on the Barrow Environmental Observatory (BEO) and the North Slope near Utqiagvik (Barrow), Alaska and 2) multiple areas on the discontinuous permafrost region of the Seward Peninsula north of Nome, Alaska.Through observations, experiments, and synthesis with existing datasets, NGEE Arctic provided an enhanced knowledge base for multi-scale modeling and contributed to improved process representation at global pan-Arctic scales within the Department of Energy's Earth system Model (the Energy Exascale Earth System Model, or E3SM), and specifically within the E3SM Land Model component (ELM).

54 ENVIRONMENTAL SCIENCES↗

From disorganized data to emergent dynamic models: Questionnaires to partial differential equations

Starting with sets of disorganized observations of spatially varying and temporally evolving systems, obtained at different (also disorganized) sets of parameters, we demonstrate the data-driven derivation of parameter dependent, evolutionary partial differential equation (PDE) models capable of generating the data. This tensor type of data is reminiscent of shuffled (multidimensional) puzzle tiles. The independent variables for the evolution equations (their “space” and “time”) as well as their effective parameters are all emergent , i.e. determined in a data-driven way from our disorganized observations of behavior in them. We use a diffusion map based questionnaire approach to build a smooth parametrization of our emergent space/time/parameter space for the data. This approach iteratively processes the data by successively observing them on the “space,” the “time” and the “parameter” axes of a tensor. Once the data become organized, we use machine learning (here, neural networks) to approximate the operators governing the evolution equations in this emergent space. Our illustrative examples are based (i) on a simple advection–diffusion model; (ii) on a previously developed vertex-plus-signaling model of Drosophila embryonic development; and (iii) on two complex dynamic network models (one neuronal and one coupled oscillator model) for which no obvious smooth embedding geometry is known a priori. This allows us to discuss features of the process like symmetry breaking, translational invariance, and autonomousness of the emergent PDE model, as well as its interpretability.

generative models↗