Search NASA⌕ Search

SEARCH · Search NASA

Results for “functional data analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 163 records · Page 9

Probing substrate water access through the O1 channel of Photosystem II by single site mutations and membrane inlet mass spectrometry

Abstract Light-driven water oxidation by photosystem II sustains life on Earth by providing the electrons and protons for the reduction of CO 2 to carbohydrates and the molecular oxygen we breathe. The inorganic core of the oxygen evolving complex is made of the earth-abundant elements manganese, calcium and oxygen (Mn 4 CaO 5 cluster), and is situated in a binding pocket that is connected to the aqueous surrounding via water-filled channels that allow water intake and proton egress. Recent serial crystallography and infrared spectroscopy studies performed with PSII isolated fromThermosynechococcus vestitus(T. vestitus) support that one of these channels, the O1 channel, facilitates water access to the Mn 4 CaO 5 cluster during its S 2 →S 3 and S 3 →S 4 →S 0 state transitions, while a subsequent CryoEM study concluded that this channel is blocked in the cyanobacteriumSynechocystis sp.PCC 6803, questioning the role of the O1 channel in water delivery. Employing site-directed mutagenesis we modified the two O1 channel bottleneck residues D1-E329 and CP43-V410 (T. vestitusnumbering) and probed water access and substrate exchange via time resolved membrane inlet mass spectrometry. Our data demonstrates that water reaches the Mn 4 CaO 5 cluster via the O1 channel in both wildtype and mutant PSII. In addition, the detailed analysis provides functional insight into the intricate protein-water-cofactor network near the Mn 4 CaO 5 cluster that includes the pentameric, near planar ‘water wheel’ of the O1 channel.

Plant Sciences↗

Modeling inclusive electron-nucleus scattering with Bayesian artificial neural networks

We introduce a Bayesian protocol based on artificial neural networks that is suitable for modeling inclusive electron-nucleus scattering on a variety of nuclear targets with quantified uncertainties. Unlike previous applications in the field, which directly parameterize the cross sections, our approach employs artificial neural networks to represent the longitudinal and transverse response functions. In contrast to cross sections, which depend on the incoming energy, scattering angle, and energy transfer, the response functions are determined solely by the energy and momentum transfer to the system, allowing the angular component to be treated analytically. We assess the accuracy and predictive power of our framework against the extensive data in the quasielastic inclusive electron-scattering database. Additionally, we present novel extractions of the longitudinal and transverse response functions and compare them with previous experimental analysis and nuclear ab-initio calculations.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

PNNL-Predictive-Phenomics/ProteoMeter

ProteoMeter is a Python package that assists in the statistical analysis of global proteomics, protein post-translation modification (PTM), and limited proteolysis (LiP) data. It contains batch correction, normalization, and statistical testing methods, as well as functions that "roll up" peptide-level data to the single-site level. It has a robust user configuration system, allowing it to flexibly integrate different types of experiment designs. For basic usage, a simple configuration file provides the essential functionality. Advanced users have access to the entire statistical pipeline for fine-tuning analyses. Processed data is easily exported to many common spreadsheet and data-frame formats.

Rozum, Jordan [Pacific Northwest National Lab]↗

Systematic uncertainty of off-shell corrections and higher-twist contribution in DIS at large x

We study the systematic uncertainty and biases introduced by theoretical assumptions needed to include large-x DIS data in a global QCD analysis. Working in the CTEQ-JLab framework, we focus on different implementations of higher-twist corrections to the nucleon structure functions and of offshell PDF deformations in deuteron targets and discuss how their interplay impacts the extraction of the d -quark PDF and the calculation of the neutron structure function at large x.

Cerutti, Matteo↗

First Detection of the Baryon Acoustic Oscillation (BAO) Feature in the 3-Point Correlation Function of DESI DR1 Luminous Red Galaxies

We present the first detection of the 3-Point Correlation Function (3PCF) Baryon Acoustic Oscillation (BAO) signal from the DESI Data Release 1 (DR1) sample of Luminous Red Galaxies (LRGs), which contains over 2.1 million galaxies. Our analysis is based on a tree-level redshift-space bispectrum template, which is then transformed to position space using the Fast Fourier Transform on Logarithmic scales (FFTLog) algorithm. We detect the BAO feature with a significance of approximately $8.1σ$ using the EZmock covariance matrix and $8.5σ$ using the analytical covariance matrix, for the full LRG redshift range ($0.4

Kamalinejad, Farshad [Florida U.] (ORCID:000000017↗

Quantitative Imaging of Cobalt Phthalocyanine Distribution on Carbon Nanotubes: A Deep Learning Approach to Catalyst Characterization

Electrochemical reduction of carbon dioxide (CO 2 ) offers a pathway to valuable products, with catalysts playing a crucial role. This study investigates the distribution of cobalt tetraaminophthalocyanine (CoPc-NH 2 ) immobilized on carbon nanotubes (CNTs), utilizing high-angle annular dark-field scanning transmission electron microscopy (HAADF-STEM) to characterize CoPc-NH 2 distribution. A challenge in the quantitative HAADF-STEM analysis is the introduction of bias from manual Co atom identification. To address this, we developed and trained a convolutional neural network (CNN) using a data set generated from images of CoPc-NH 2 /CNT samples with varying Co loadings. The CNN, implemented in TensorFlow and Keras, facilitated Co atom detections. Analysis of the CNN-generated data confirmed a correlation between Co loading and surface density, consistent with findings from UV–vis spectroscopy. Furthermore, the application of Ripley’s L(d) function highlighted the presence of slight Co atom clustering. Furthermore, this work demonstrates the utility of the combined HAADF-STEM and CNN approach for providing spatially resolved information about catalyst distribution on nonplanar supports, revealing structural details that are typically lost through other characterization methods.

HAADF-STEM↗

Feedback Controllability Components Analysis (FCCA) v1.0

FCCA is a linear dimensionality reduction method that find subspaces of high-dimensional time-series data that are most feedback controllable. The key innovation is to formulate an objective function that quantifies the joint cost of state reconstruction and state regulation that can be evaluated from purely observational data. To do this, it leverages the duality between controllability and observability. We provide analytic results demonstrating the validity of the cost function. We evaluated this method in both synthetic and real neural data (from multiple organisms and brain areas).

Kumar, Ankit↗

From Reads to Function Workshop - Milano 2026

The Bicocca Sampling Days (BSDs) model offers a reproducible “citizen science” framework integrating research, education, and public engagement through large-scale microbiome sampling, followed by a workshop of data analysis on select samples. We identified 9 bacterial and archaeal metagenome-assembled genomes from six soil samples across three separate sampling days in two approaches with indidivual sample and replicate co-assembly spanning three unique classes, providing genomic insights into microbial nutrient cycling in these systems.

59 BASIC BIOLOGICAL SCIENCES↗

INSPIRED: Inelastic neutron scattering prediction for instantaneous results and experimental design

Inelastic neutron scattering (INS) has unique advantages in probing how atoms vibrate and how the vibrations propagate and interact. Such dynamic information is crucial in understanding various material properties, from heat capacity, thermal conductivity, phase transitions, and chemical reactions to more exotic quantum behavior. The analysis and interpretation of the INS spectra often start from a model structure of the sample, followed by a series of calculations to obtain the simulated spectra to compare with experiments. The conventional way to perform such calculations usually requires significant time, computing resources, and specialized expertise. Here, we present a new program named INSPIRED (Inelastic Neutron Scattering Prediction for Instantaneous Results and Experimental Design), which enables users to perform rapid INS simulations in several different ways on their personal computers in just a few clicks, with the crystal structure as the only input file. Specifically, the users can choose a pre-trained symmetry-aware neural network (coupled with an autoencoder) to predict the phonon density of states (DOS), 1D S(E) and 2D S(|Q|,E) spectra for any given structure. One can also choose an existing density functional theory (DFT) calculation from a database (containing over 12,000 crystals), and quickly obtain the simulated INS spectra for single crystals and powders. It is also possible to use pre-trained universal machine learning force fields to relax a given crystal structure, calculate the phonon dispersion and DOS, and, subsequently, the INS spectra. All these functions are implemented with a PyQt graphic user interface. Finally, we expect these new tools will benefit broad user communities and significantly improve the efficiency of experiment design, execution, and data analysis for INS.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Accuracy of kinetic equilibrium reconstruction of NSTX and NSTX-U plasmas and its impact on the transport and stability analysis

An accurate magnetohydrodynamic (MHD) equilibrium reconstruction is an essential starting point for stability and transport plasma analysis. Herein this work describes an approach for obtaining kinetic equilibrium reconstructions using the OMFIT framework, which has been applied for the first time to spherical tokamak data from NSTX and NSTX-U. The EFIT equilibrium solver is integrated with experimental data analysis procedures and subsequent TRANSP transport simulations to enhance the accuracy of the reconstruction, in particular, at the edge region, by adding constraints on the total pressure and current density profiles, based on the transport code solution. The accuracy of the equilibrium reconstruction depends on the uncertainty and number of constraints, as well as the choice of basis functions to represent the pressure and current density profiles. Improved fidelity of the equilibrium reconstruction is demonstrated by reducing the variability of the magnetic axis and boundary locations from several centimeters, for reconstructions based on magnetic and experimental pressure constraints, to only several millimeters, for kinetic reconstructions based on transport code constraints, when different representations of basis functions were tested. The variability of the safety factor on axis was reduced ten times in the same sensitivity study. The accuracy of the equilibrium reconstruction and subsequent mapping of the experimental kinetic profile data have a significant impact on the trapped gyro Landau fluid and linear CGYRO turbulence simulations, which predict different spectra of unstable modes and turbulent fluxes for cases with different numbers of constraints in the equilibrium reconstruction. Conversely, the stability analysis performed using the GATO code shows plasmas that are stable to n = 1 MHD modes in both equilibria using magnetic and experimental pressure constraints as well as the transport code constrained equilibrium. However, a scan of parameters away from these conditions shows considerable deviation in the threshold of unstable modes between these reconstructions. Therefore, for reliable plasma analysis and use in turbulence and stability calculations, a high-fidelity equilibrium reconstruction with accurate kinetic constraints based on transport code solutions is necessary.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Chemoproteogenomic stratification of the missense variant cysteinome

Abstract Cancer genomes are rife with genetic variants; one key outcome of this variation is widespread gain-of-cysteine mutations. These acquired cysteines can be both driver mutations and sites targeted by precision therapies. However, despite their ubiquity, nearly all acquired cysteines remain unidentified via chemoproteomics; identification is a critical step to enable functional analysis, including assessment of potential druggability and susceptibility to oxidation. Here, we pair cysteine chemoproteomics—a technique that enables proteome-wide pinpointing of functional, redox sensitive, and potentially druggable residues—with genomics to reveal the hidden landscape of cysteine genetic variation. Our chemoproteogenomics platform integrates chemoproteomic, whole exome, and RNA-seq data, with a customized two-stage false discovery rate (FDR) error controlled proteomic search, which is further enhanced with a user-friendly FragPipe interface. Chemoproteogenomics analysis reveals that cysteine acquisition is a ubiquitous feature of both healthy and cancer genomes that is further elevated in the context of decreased DNA repair. Reference cysteines proximal to missense variants are also found to be pervasive, supporting heretofore untapped opportunities for variant-specific chemical probe development campaigns. As chemoproteogenomics is further distinguished by sample-matched combinatorial variant databases and is compatible with redox proteomics and small molecule screening, we expect widespread utility in guiding proteoform-specific biology and therapeutic discovery.

Desai, Heta (ORCID:0000000343621707)↗

Population‐level gene expression can repeatedly link genes to functions in maize

SUMMARY Transcriptome‐wide association studies (TWAS) can provide single gene resolution for candidate genes in plants, complementing genome‐wide association studies (GWAS) but efforts in plants have been met with, at best, mixed success. We generated expression data from 693 maize genotypes, measured in a common field experiment, sampled over a 2‐h period to minimize diurnal and environmental effects, using full‐length RNA‐seq to maximize the accurate estimation of transcript abundance. TWAS could identify roughly 10 times as many genes likely to play a role in flowering time regulation as GWAS conducted data from the same experiment. TWAS using mature leaf tissue identified known true‐positive flowering time genes known to act in the shoot apical meristem, and trait data from a new environment enabled the identification of additional flowering time genes without the need for new expression data. eQTL analysis of TWAS‐tagged genes identified at least one additional known maize flowering time gene through trans ‐eQTL interactions. Collectively these results suggest the gene expression resource described here can link genes to functions across different plant phenotypes expressed in a range of tissues and scored in different experiments.

Torres‐Rodríguez, J. Vladimir↗

A Science Gateway for the Repeatable Analysis of Machine Learning Predicted Gravity Anomalies

In recent years, deep learning has become an increasingly popular alternative for modeling in geoscience applications due to its scalability and efficiency. However, the interpretability, compute, data volume, and hyperparameter tuning requirements of deep learning models make development and monitoring difficult. Furthermore, model explainability and communicating results obtained by these models to users or domain experts is a challenge, as domain experts in geoscience also need to have a deep understanding of how those models function in order to support their scientific works. Here, we describe a science gateway and machine learning pipeline for predicting gravity anomalies from geophysical data. The gateway, built on open-source technologies, provides a holistic view of the pipeline through interactive visualizations aimed at enabling efficient exploratory data analysis. The repeatability, reproducibility, and monitoring capabilities of this overall system allow us to iterate and analyze at scale. Using this pipeline and gateway, we can repeatedly produce accurate high-resolution gravity anomaly datasets. By describing the underlying technologies, implementation, and results, here we provide a foundation for the broader adoption of science gateways into cross-cutting geoscience and machine learning research projects as a means to improve the scientific discovery and collaboration in the geophysics and computational sciences community.

58 GEOSCIENCES↗

RWRtoolkit: multi-omic network analysis using random walks on multiplex networks in any species

Abstract We introduce RWRtoolkit, a multiplex generation, exploration, and statistical package built for R and command-line users. RWRtoolkit enables the efficient exploration of large and highly complex biological networks generated from custom experimental data and/or from publicly available datasets, and is species agnostic. A range of functions can be used to find topological distances between biological entities, determine relationships within sets of interest, search for topological context around sets of interest, and statistically evaluate the strength of relationships within and between sets. The command-line interface is designed for parallelization on high-performance cluster systems, which enables high-throughput analysis such as permutation testing. Several tools in the package have also been made available for use in reproducible workflows via the KBase web application.

Kainer, David (ORCID:0000000172714676)↗

The landscape of regulatory element evolution in a C4 perennial grass

Gene regulatory evolution is a well-known source of phenotypic diversity and adaptive evolution. Although cis-regulatory elements (CREs) play a vital role in gene expression evolution, the molecular evolution of CREs remains mostly unknown due to the difficulty in identifying and characterizing these functional elements. Comparative genomic analyses of noncoding DNA can be leveraged to identify conserved noncoding sequences (CNS), many of which may harbor functional CREs conserved by purifying selection. However, purely computational inference of CREs from putative CNS can be erroneous due to the complex genomic architecture in plants. One promising experimental approach to identify CREs is by profiling accessible chromatin regions (ACRs) that are often associated with the location of CREs. In this study, we use comparative genomics along with the profiling of ACRs to study the molecular evolution of putative functional noncoding regulatory regions in Panicoid grasses. We identified sets of CNS that varied in relationship to the degree of evolutionary divergence among the studied taxa, including identifying core-Panicoid-CNS. We augmented this analysis by profiling ACRs in Panicum hallii ecotypes using ATAC-seq. ACRs had low SNP density at the summit, harbored a high frequency of core-Panicoid-CNS, and were enriched with expression QTL. These data help to annotate the P. hallii genome for putative functional elements and suggest that a large proportion of these ACRs are evolving under purifying selection. Turnover in CNS and ACR between ecotypes of P. hallii identifies a small set of putatively divergent CREs that may underlie differences in gene regulation between genotypes from inland and coastal habitats. In summary, we profiled ACRs in Panicoid grasses and integrated this data with our putative CNS prediction framework, which provides unique insight into patterns of polymorphism and divergence in CREs in C4 perennial grasses.

59 BASIC BIOLOGICAL SCIENCES↗

Constructing Data-Driven Predictions at the Far Detector for NOvA's Neutrino Oscillation Analysis.

NOvA, is a two-detector, long-baseline neutrino oscillation experiment located at Fermilab, Batavia, IL, USA. It is designed primarily to constrain neutrino oscillation parameters using $\nu_\mu \ (\bar{\nu}_\mu)$ disappearance and $\nu_e \ (\bar{\nu}_e)$ appearance data. The Neutrinos at Main Injector (NuMI) beamline at Fermilab provides a high purity 900 KW intense beam of neutrinos and anti-neutrinos to NOvA. The NOvA Near Detector, located 100m underground and 1km away from the beam source, observes the un-oscillated $\nu_\mu \ (\bar{\nu}_\mu)$ and beam $\nu_e \ (\bar{\nu}_e)$ event spectrum. The Far Detector, located in Ash River, MN, USA, is 809 km from the ND and records the oscillated $\nu_e \ (\bar{\nu}_e)$ and the un-oscillated $\nu_\mu \ (\bar{\nu}_\mu)$ event spectrum. NOvA uses a data-driven technique called extrapolation to predict the expected number of $\nu_\mu \ (\bar{\nu}_\mu)$ and $\nu_e \ (\bar{\nu}_e)$ events at the Far Detector using the Near Detector data. The use of data from a functionally equivalent Near Detector provides a powerful constraint on the systematic uncertainties in NOvA neutrino oscillation analyses. As NOvA continues to add data statistics, a robust constraint on systematics becomes more crucial for neutrino oscillation analysis. The details of the NOvA neutrino oscillation analysis framework and how it constrains dominant systematic uncertainties using the Near Detector data will be discussed in this poster.

43 PARTICLE ACCELERATORS↗

Glauber-Theory Calculations of High-Energy Nuclear Scattering Observables Using Variational Monte Carlo Wave Functions

Experiments using intermediate- to high-energy radioactive nuclear beams present numerous findings. Extracting important properties of physical observables relies on a firm theoretical analysis. Though Glauber theory is believed to work well, no convincing calculation has so far been done. Here, we perform ab initio Glauber theory calculations of both elastic differential cross sections and total reaction cross sections for p+ 12 C, 12 C+ 12 C, and 6 He+ 12 C systems. The wave functions of both 6 He and 12 C are generated by variational Monte Carlo calculations with spatial and spin-isospin correlations induced by realistic two- and three-nucleon potentials. Glauber’s phase-shift function is computed by Monte Carlo integration up to all orders of nucleon-nucleon multiple scatterings. We show an excellent performance of the Glauber description to the selected data on the above systems. We also find that the cumulant expansion of the phase-shift function converges rapidly up to the second order for the above systems. This finding will open up interesting applications for the analysis of high-energy nuclear experiments.

Horiuchi, W. [Osaka Metropolitan University (Japan↗

Analysis of Tar and Oil Derived from Pyrolysis and Copyrolysis of Waste Plastics and Biomass

Pyrolysis has been proposed as a potential technology for managing the growing volume of plastic waste generated worldwide. Co-pyrolysis of plastic waste with biomass is a promising technology for generating fuel and chemical products. However, this process generates tar as a waste product. The chemical properties of this tar have yet to be thoroughly analyzed. Further, this study presents the results of gas chromatography–mass spectrometry (GC–MS), Fourier-transform infrared spectroscopy (FTIR), and thermogravimetric analysis (TGA) of oil and tar obtained from the pyrolysis of pure plastics including high-density polyethylene (HDPE), low-density polyethylene (LDPE), polyethylene (PE), polystyrene (PS), and plastic-biomass mixtures. GC–MS analysis revealed the presence of C 7 –C 37 carbon-containing hydrocarbons, which include alkanes and alkenes as the dominant products. FTIR data revealed the presence of various functional groups, including alcohols, aldehydes, ketones, and carboxylic acids, indicating the complexity of the pyrolysis and copyrolysis oil obtained from waste plastics and biomass. TGA data show that tar from all four plastics has a higher decomposition rate, suggesting the presence of heavier hydrocarbons compared with their corresponding oils. This research will be of interest to researchers looking to advance the study of plastic and biomass waste management.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗