Search NASA⌕ Search

SEARCH · Search NASA

Results for “Data Inference”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 451 records · Page 25

Protocol for applying a network-enabled gene discovery pipeline to non-model plant species

Identifying upstream regulators of key genes is essential for understanding gene regulatory mechanisms and translating these insights into functional targets. Here, we present a protocol for applying the network-enabled gene discovery pipeline (NEEDLE) to non-model plant species. We describe steps for environment setup, data preparation, computational analysis, expected outputs, and parameter considerations. NEEDLE integrates RNA sequencing (RNA-seq) processing, weighted gene co-expression analysis (WGCNA), Gene Network Inference with Ensemble of trees (GENIE3), and promoter conservation analysis to prioritize candidate transcriptional regulators.

Plant Sciences↗

Multi-fidelity equations of state and transport coefficient datasets for pulsed-power applications

Reliably simulating experiments relevant to the National Nuclear Security Administration (NNSA) requires a detailed description of material properties across a wide range of conditions. Such properties include the equations of state, charged-particle transport coefficients, and optical properties like the opacity. Together, these properties make up the material models used in radiation-magnetohydrodynamic simulations of nuclear fusion experiments. Many of these models do not incorporate uncertainties in the data used to produce them. It is unknown whether these uncertainties significantly impact the interpretation of simulation results and diagnostics. The purpose of this work is to quantify how such uncertainties impact simulations of pulsed-power experiments. We accomplished this task by first assessing discrepancies between approaches used to generate the data. This included bringing together members of the high-energy-density community spanning the three NNSA laboratories and multiple universities. Then, using these data, we developed a general framework that systematically incorporates physical uncertainties within the material models suitable for uncertainty quantification analyses. The framework utilizes machine learning, Bayesian inference, and incorporates multi-fidelity datasets. We demonstrated the framework by quantifying the impact that material model uncertainties have on simulations of pulsed-power experiments underway on Z at Sandia National Laboratories. As a result of this work, we discovered that modest uncertainties in material models (roughly 20%) correspond to significant uncertainties in the outputs from simulations. Our framework has enabled rapid construction of material models through an automated procedure and allows for the generation of material models of interest to the NNSA.

36 MATERIALS SCIENCE↗

Differentially Private Adaptive Noise Injection (DP-ANI) v1.0

Location data is collected from users continuously to understand their mobility patterns. Releasing the user trajectories may compromise user privacy. Therefore, the general practice is to release aggregated location datasets. However, private information may still be inferred from an aggregated version of location trajectories. Differential privacy (DP) protects the query output against inference attacks regardless of background knowledge. This software implements a differential privacy-based privacy model that protects the user's origins and destinations from being inferred from aggregated mobility datasets. This is achieved by injecting Planar Laplace noise to the user origin and destination GPS points. The noisy GPS points are then transformed into a link representation using a link-matching algorithm. Finally, the link trajectories form an aggregated mobility network. The injected noise level is selected using the Sparse Vector Mechanism. This DP selection mechanism considers the link density of the location and the functional category of the localized links. Compared to the different baseline models, including a k-anonymity method, our differential privacy-based aggregation model offers query responses that are close to the raw data in terms of aggregate statistics at both the network and trajectory-levels with maximum 9% deviation from the baseline in terms of network length.

Peisert, Sean [Lawrence Berkeley National Laborato↗

OrthoPhyl—streamlining large-scale, orthology-based phylogenomic studies of bacteria at broad evolutionary scales

Abstract There are a staggering number of publicly available bacterial genome sequences (at writing, 2.0 million assemblies in NCBI's GenBank alone), and the deposition rate continues to increase. This wealth of data begs for phylogenetic analyses to place these sequences within an evolutionary context. A phylogenetic placement not only aids in taxonomic classification but informs the evolution of novel phenotypes, targets of selection, and horizontal gene transfer. Building trees from multi-gene codon alignments is a laborious task that requires bioinformatic expertise, rigorous curation of orthologs, and heavy computation. Compounding the problem is the lack of tools that can streamline these processes for building trees from large-scale genomic data. Here we present OrthoPhyl, which takes bacterial genome assemblies and reconstructs trees from whole genome codon alignments. The analysis pipeline can analyze an arbitrarily large number of input genomes (>1200 tested here) by identifying a diversity-spanning subset of assemblies and using these genomes to build gene models to infer orthologs in the full dataset. To illustrate the versatility of OrthoPhyl, we show three use cases: E. coli/Shigella, Brucella/Ochrobactrum and the order Rickettsiales. We compare trees generated with OrthoPhyl to trees generated with kSNP3 and GToTree along with published trees using alternative methods. We show that OrthoPhyl trees are consistent with other methods while incorporating more data, allowing for greater numbers of input genomes, and more flexibility of analysis.

59 BASIC BIOLOGICAL SCIENCES↗

Inverse prediction of PuO2 processing conditions using Bayesian seemingly unrelated regression with functional data

Over the past decade, a variety of innovative methodologies have been developed to better characterize the relationships between processing conditions and the physical, morphological, and chemical features of special nuclear material (SNM). Different processing conditions generate SNM products with different features, which are known as “signatures” because they are indicative of the processing conditions used to produce the material. These signatures can potentially allow a forensic analyst to determine which processes were used to produce the SNM and make inferences about where the material originated. This article investigates a statistical technique for relating processing conditions to the morphological features of PuO 2 particles. We develop a Bayesian implementation of seemingly unrelated regression (SUR) to inverse-predict unknown PuO 2 processing conditions from known PuO 2 features. Model results from simulated data demonstrate the usefulness of the technique. Applied to empirical data from a bench-scale experiment specifically designed with inverse prediction in mind, our model successfully predicts nitric acid concentration, while results for Pu concentration and precipitation temperature were equivalent to a simple mean model. Our technique compliments other recent methodologies developed for forensic analysis of nuclear material and can be generalized across the field of chemometrics for application to other materials.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Structure of Self-Generated Magnetic Fields in Laser-Solid Interaction from Proton Tomography

Self-generated magnetic fields in laser-solid interactions are experimentally characterized to reveal the 3D location and local field strength, rather than path-integrated quantities, using multi-view proton radiography and tomographic inversion. We infer magnetic fields that extend several millimeters off the target into the hot, rarefied corona, sufficient to strongly magnetize the plasma (Ω e τ e ≫ 1). The data are compared to MHD simulations incorporating recent improvements in modeling magnetic field generation and transport; the volume-averaged coronal field strength and magnetic flux agree to within 25% using a model with magnetic re-localization of transport, although the near-target morphology is not reproduced. This work demonstrates tomographic proton radiography as a valuable tool for investigating magnetic fields in laser-produced plasmas.

High-energy-density plasmas↗

Representing Complex Systems as Graphs for Debugging and Predictive Maintenance-Preliminary Thoughts

Representing complex systems as graphs enables use of mathematical tools to identify faults or predict failures. Graph nodes correspond to individual modules or subsystems, and edges link coupled system parts. ‘Probes’ measure the node outputs, monitoring the system health for unexpected behavior. Assuming one cannot probe every point, within a system, the fault correlates to a region—not necessarily the specific location. Bayesian networks trained to understand fault patterns can accurately identify the source. The diagnostic tool described aides debugging by pinpointing system failure causes. For predictive maintenance, probe data develop probability distribution functions describing subsystem mean time to failure. Unit lifetime can be estimated through these probability distributions. Two approaches include using Bayesian classifiers to infer the system failure source and developing maintenance schedules by treating systems as collections of random variables. When failure behavior does not follow a closed form function, use of similarity models is proposed.

97 MATHEMATICS AND COMPUTING↗

Simultaneous global and local clustering in multiplex networks with covariate information

Understanding both global and layer-specific group structures is useful for uncovering complex patterns in networks with multiple interaction types. In this work, we introduce a new model, the hierarchical multiplex stochastic blockmodel, which simultaneously detects communities within individual layers of a multiplex network while inferring a global node clustering across the layers. A stochastic blockmodel is assumed in each layer, with probabilities of layer-level group memberships determined by a node’s global group assignment. Our model uses a Bayesian framework, employing a probit stick-breaking process to construct node-specific mixing proportions over a set of shared Griffiths–Engen–McCloseky distributions. These proportions determine layer-level community assignment, allowing for an unknown and varying number of groups across layers, while incorporating nodal covariate information to inform the global clustering. We propose a scalable variational inference procedure with parallelisable updates for application to large networks. Extensive simulation studies demonstrate our model’s ability to accurately recover both global and layer-level clusters in complicated settings, and applications to real data showcase the model’s effectiveness in uncovering interesting latent network structure.

community detection↗

Analysis of Polarized Dust Emission Using Data from the First Flight of SPIDER

Using data from the first flight of Spider and from the Planck High Frequency Instrument, we probe the properties of polarized emission from interstellar dust in the Spider observing region. Component-separation algorithms operating in both the spatial and harmonic domains are applied to probe their consistency and to quantify modeling errors associated with their assumptions. Analyses of diffuse Galactic dust emission spanning the full Spider region demonstrate (i) a spectral energy distribution that is broadly consistent with a modified-blackbody (MBB) model with a spectral index of β d = 1.45 ± 0.05 (1.47 ± 0.06) for E (B)-mode polarization, slightly lower than that reported by Planck for the full sky; (ii) an angular power spectrum broadly consistent with a power law; and (iii) no significant detection of line-of-sight polarization decorrelation. Tests of several modeling uncertainties find only a modest impact (~10% in σ r ) on Spider's sensitivity to the cosmological tensor-to-scalar ratio. The size of the Spider region further allows for a statistically meaningful analysis of the variation in foreground properties within it. Assuming a fixed dust temperature T d = 19.6 K, an analysis of two independent subregions of that field results in inferred values of β d = 1.52 ± 0.06 and β d = 1.09 ± 0.09, which are inconsistent at the 3.9σ level. Furthermore, a joint analysis of Spider and Planck 217 and 353 GHz data within one subregion is inconsistent with a simple MBB at more than 3σ, assuming a common morphology of polarized dust emission over the full range of frequencies. This evidence of variation may inform the component-separation approaches of future cosmic microwave background polarization experiments.

79 ASTRONOMY AND ASTROPHYSICS↗

Probing cosmic velocities with the pairwise kinematic Sunyaev-Zel’dovich signal in DESI Bright Galaxy Sample DR1 and ACT DR6

We present a measurement of the pairwise kinematic Sunyaev-Zel’dovich (kSZ) signal using the Dark Energy Spectroscopic Instrument (DESI) Bright Galaxy Sample (BGS) Data Release 1 (DR1) galaxy sample overlapping with the Atacama Cosmology Telescope (ACT) CMB temperature map. Our analysis makes use of 1.6 million galaxies with stellar masses log⁡ 𝑀 ⋆ /𝑀 ⊙ >10, and we explore measurements across a range of aperture sizes (2.1′ <𝜃 ap <3.5′) and stellar mass selections. This statistic directly probes the velocity field of the large-scale structure, a unique observable of cosmic dynamics and modified gravity. In particular, at low redshifts, this quantity is especially interesting, as deviations from General Relativity are expected to be largest. Notably, our result represents the highest-significance low-redshift (𝑧 ∼ 0.3) detection of the kSZ pairwise effect yet. In our most optimal configuration (𝜃 ap =3.3′, log⁡ 𝑀 ⋆ >11), we achieve a 5⁢𝜎 detection. Assuming that an estimate of the optical depth and galaxy bias of the sample exists via e.g., external observables, this measurement constrains the fundamental cosmological combination 𝐻 0 ⁡𝑓⁡𝜎$^2_8$. A key challenge is the degeneracy with the galaxy optical depth. We address this by combining CMB lensing, which allows us to infer the halo mass and galaxy population properties, with hydrodynamical simulation estimates of the mean optical depth, $\bar{𝜏}$ . We stress that this is a proof-of-concept analysis; with BGS DR2 data we expect to improve the statistical precision by roughly a factor of two, paving the way toward robust tests of modified gravity with kSZ-informed velocity-field measurements at low redshift.

Hadzhiyska, Boryana [Institute of Astronomy; Kavli↗

Plasma decay of nanosecond pulsed laser-produced Ar and Ar–H 2 O sparks at atmospheric pressure

Time-resolved diagnostics were applied to investigate free-electron properties in nanosecond laser-produced discharges generated in atmospheric pressure Ar and in Ar–3%H 2 O. The discharges were generated using 23 ns, 1064 nm laser pulses. Broadband plasma imaging and laser Thomson scattering were combined with optical emission spectroscopy, with particular emphasis on the Stark broadening of the H α and H β lines. The plasma exhibited a bright emission that persists for up to 30–40 µs after breakdown. Plasma emission was then followed by a very weak glow emission that persisted for up to 19 ms after breakdown. Peak electron number density of ∼2 × 10 17 cm −3 and electron temperature of ∼7 eV were measured. An excellent agreement between both techniques was obtained regarding absolute electron number densities. The inferred free-electron temporal decay dynamics are consistent with processes dominated by hydrodynamic expansion and two- and three-body electron–ion recombination. These results provide benchmark data for modeling nanosecond laser discharges and demonstrate the reliability of combining Thomson scattering with Stark broadening in atmospheric laser sparks.

Thomson scattering↗

Selecting samples of galaxies with fewer Fingers-of-God

The radial positions of galaxies inferred from their measured redshift appear distorted due to their peculiar velocities. We argue that the contribution from stochastic velocities — which gives rise to `Fingers-of-God' (FoG) anisotropy in the inferred maps — does not lend itself to perturbative modelling already on scales targeted by current experiments. To get around this limitation, we propose to remove FoG using data-driven indicators of their abundance that are local in nature and thus avoid selection biases. In particular, we show that the scale where the measured power spectrum quadrupole changes sign is tightly anti-correlated with both the satellite fraction and the velocity dispersion, and can thus be used to select galaxy samples with fewer FoG. In addition, we show that excluding galaxies in haloes more massive than a given mass threshold can help to discard many of the most problematic galaxies. Such selection could be achieved in practice using maps of the thermal Sunyaev-Zel'dovich distortion of the cosmic microwave background frequency spectrum. These techniques could potentially improve reconstructions of the large-scale velocity and displacement fields from the redshift-space positions of galaxies. They may also extend the reach of perturbative models for galaxy clustering, though in practice we find only marginal gains when fitting one-loop EFTofLSS models to simulations with mitigated FoG due to the relevance of other effects entering at two-loop order.

cosmological parameters from LSS↗

Multiclass Classification Using Bayesian Multivariate Adaptive Regression Splines

We present a new Bayesian model for the problem of multiclass classification. In this model, the probabilities of class membership of a given observation are determined by the mean of a latent Gaussian distribution. The mean functions of this latent distribution consist of combinations of highly flexible basis functions of the inputs: multivariate adaptive regression splines (MARS), first developed for multiple regression. We use reversible jump Markov chain Monte Carlo to make inference on the classification model, including the number of basis functions. We compare the probabilistic classification performance of our proposed approach to existing methods on simulated and benchmark data, and compare uncertainty estimates on simulated data. Our proposed method compares favorably with existing Bayesian and frequentist multiclass classification methods in out-of-sample probabilistic classification, and uncertainty estimation of these probabilistic classifications. We examine the fit of the proposed method to a data set of hurricane storm surge levels near Delaware Bay, US, and conclude that sea level rise is a key contributor to damage delivered by storm surge.

97 MATHEMATICS AND COMPUTING↗

Robust Measurement of Stellar Streams around the Milky Way: Correcting Spatially Variable Observational Selection Effects in Optical Imaging Surveys

Observations of density variations in stellar streams are a promising probe of low-mass dark matter substructure in the Milky Way. However, survey systematics such as variations in seeing and sky brightness can also induce artificial fluctuations in the observed densities of known stellar streams. These variations arise because survey conditions affect both object detection and star–galaxy misclassification rates. To mitigate these effects, we use Balrog synthetic source injections in the Dark Energy Survey (DES) Y3 data to calculate detection rate variations and classification rates as functions of survey properties. We show that these rates are nearly separable with respect to survey properties and can be estimated with sufficient statistics from the synthetic catalogs. Applying these corrections reduces the standard deviation of relative detection rates across the DES footprint by a factor of 5, and our corrections significantly change the inferred linear density of the Phoenix stream when including faint objects. Additionally, for artificial streams with DES-like survey properties we are able to recover density power spectra with reduced bias. We also find that uncorrected power-spectrum results for Legacy Survey of Space and Time (LSST)-like data can be around 5 times more biased, highlighting the need for such corrections in future ground-based surveys.

79 ASTRONOMY AND ASTROPHYSICS↗

Feasibility of Correlation-Aware Inference and Universal Precision Scaling in Bonse–Hart Ultra-Small-Angle Neutron Scattering

Bonse–Hart ultra-small-angle neutron scattering (USANS) provides access to micrometre-scale structure, but useful measurements often require long counting times. In this work, we test whether the expected smoothness of the scattering profile can be exploited to improve data quality at lower counting statistics. We apply a Gaussian-process-based method to Bonse–Hart USANS data and evaluate its performance on pseudo-measurements generated from high-statistics experiments under Poisson statistics. This provides a stringent test of how well the underlying I(Q) profile can be reconstructed when the available counts are substantially reduced. We further show that, in the counting-limited regime, the reconstruction error follows a universal scaling behaviour that differs from the usual independent-counting expectation. At higher counts, the improvement crosses over to a resolution-limited regime set by analyser-angle discretization and rocking-curve width. These results clarify when correlation-aware inference is useful in USANS and provide a practical basis for improving measurement efficiency and beam-time usage.

Tung, Chi-Huan [ORNL] (ORCID:0000000221972074)↗

Earth System Reanalysis in Support of Climate Model Improvements

Recent climate model developments, established through increased model resolution, have led to substantial improvements in model simulations of the time-evolving, coupled Earth system and its subcomponents. However, regardless of resolution, climate models will always produce climate features and variability that differ from the real world and will be prone to biases. This is due to many remaining uncertainties, such as in parametric and structural model uncertainty, in the initial conditions prescribed, and in the prescribed (scenario) forcing which varies on decadal to centennial timescales. Further model improvements are expected to arise specifically from improved representation of physical processes realized through model-data fusion. This will create an unprecedented opportunity to better exploit a large array of Earth observations, from in situ measurements to weather radars and satellite observations, as the resolved scales of the models approach those of the observations. For this, climate DA will be the central tool to bring models and observations into consistency, by improving initial conditions, inferring uncertain model parameters and structure, and quantifying uncertainty. Generally, there will be advantages and complementarities of adjoint-based smoother approaches, ensemble-based filter approaches, or new ML-inspired approaches. Yet, the ever-increasing model resolution will present growing challenges arising from computational cost, calling for new ways of performing data assimilation and model optimization. Using the complementarity in a hybrid approach, blending tools and concepts from variational, ensemble and ML methods might be what is required in the future. In this context ML could be important to handle non-linear responses, and to better approximate non-Gaussian distributions.

54 ENVIRONMENTAL SCIENCES↗

Sequential spectral line analysis for accurate density and temperature diagnosis of laboratory opacity measurements

The accuracy of iron opacity calculated in stellar interiors has been questioned since the discovery of the “solar problem” and the discrepancies between the measured and modeled iron opacity reported in 2015. Experimental opacity benchmarks require accurate temperature and density measurements, which were inferred by analyzing tracer magnesium spectra in those experiments. Could the observed discrepancy be explained by insufficient accuracy in the inferred temperature, density, and their uncertainties? Previous analyses may have yielded biased results due to three limitations: (1) simultaneous multi-line fitting, (2) approximations in line-shape models, and (3) exclusion of certain spectral lines due to insufficient background characterization. Notably, the first issue is a common concern for many inversion methods, including Bayesian inferences. We present a refined analysis method that overcomes these limitations, applied to three categories of iron opacity experiments (Anchor 1, 2, and 3). In particular, the sequential fitting method yields unbiased results with more realistic uncertainties by accounting for line inconsistencies in the parameter uncertainties. The average electron temperature and density values are 162 ± 6 eV and (7.0 ± 1.9) × 10 21 cm −3 for six Anchor 1 experiments, 189 ± 7 eV and (3.4 ± 0.3) × 10 22 cm −3 for 21 Anchor 2 experiments, and 201 ± 6 eV and (4.8 ± 1.1) × 10 22 cm −3 for nine Anchor 3 experiments. These results show ∼4% temperature and ∼20% density reproducibility over a decade, which also aligns with the inferred parameter uncertainties. In conclusion, the resulting temperature and density uncertainties lead to a quasi-continuum iron opacity variation of ±4%–7% for wavelengths below 9.5 Å, which is insufficient to explain the significant model-data discrepancies reported in 2015.

Absorption spectroscopy↗

Radiochemical transport analysis of gamma spectroscopic data to support estimation of molten salt reactor off-gas inventories

This work introduces a novel application of radiochronometry to estimate nuclide inventories in molten salt reactor off-gas systems based on gamma spectroscopic data from the Molten Salt Reactor Experiment. By analyzing isotopic, isobaric, and isomeric activity ratios, key depletion model parameters related to species transport within the reactor system could be inferred. The findings demonstrate the potential of leveraging a limited subset of gamma spectroscopy measurements to accurately estimate nuclide inventories throughout the off-gas system. The approach can be useful in reactor design activities and support analyses relevant to operations, safety, security, and safeguards.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗