Search NASASearch

SEARCH · Search NASA

Results for “component analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Unbiased primordial gravitational wave inference from the CMB with SMICA

The detection of primordial gravitational waves in Cosmic Microwave Background B-mode polarization observations requires accurate and robust subtraction of astrophysical contamination. We show, using a blind Spectral Matching Independent Component Analysis, that it is possible to infer unbiased estimates of the primordial B-mode signal from ground-based observations of a small patch of sky even for highly complex foreground contamination. This work, originally performed in the context of configuration studies for a future CMB-S4 observatory, is highly relevant for the analysis of observations by the current generation of CMB experiments.

cosmological parameters from CMBR

A 28 nm multiply-accumulate ASIC architecture for on-chip data compression in MHz frame rate X-ray and electron pixel detectors

Modern X-ray detector systems urgently require compact, efficient, and fast data compression schemes to handle the transmission of big data from pixel arrays, enabling frame rates in the MHz regime. Here, in this work, a data compression ASIC that implements a streaming fixed-length lossy compression scheme is introduced and analyzed, proving the feasibility and benefits of on-chip compression. The compression scheme utilizes a vector matrix product logic, which performs a number of floating-point multiplications, additions, and accumulations. The logic is verified, synthesized, and shown to fit in the area resource available for the X-ray detector under study, which comprises 192 × 168 pixels each of 12-bit width, and having a total area of 20 mm× 20 mm, about 2 mm× 20 mm of which are available for the digital logic. Several system architectures, precisions, and compression ratios ranging from 100 to 250 were analyzed to pave the way for on-chip fixed-length compression (e.g., principal component analysis, singular value decomposition) and data reduction (e.g., azimuthal integration) for X-ray and electron detectors.

Data compression

Towards revealing intrinsic vortex-core states in Fe-based superconductors through statistical discovery

Abstract In type-II superconductors, electronic states within magnetic vortices hold crucial information about the paring mechanism and can reveal non-trivial topology. While scanning tunneling microscopy/spectroscopy (STM/S) is a powerful tool for imaging superconducting vortices, it is challenging to isolate the intrinsic electronic properties from extrinsic effects like subsurface defects and disorders. Here we combine STM/STS with basic machine learning to develop a method for screening out the vortices pinned by embedded disorder in iron-based superconductors. Through a principal component analysis of large STS data within vortices, we find that the vortex-core states in Ba(Fe 0.96 Ni 0.04 ) 2 As 2 start to split into two categories at certain magnetic field strengths, reflecting vortices with and without pinning by subsurface defects or disorders. Our machine-learning analysis provides an unbiased approach to reveal intrinsic vortex-core states in novel superconductors and shed light on ongoing puzzles in the possible emergence of a Majorana zero mode.

Guo, Yueming

Precise relative magnitude measurement improves fracture characterization during hydraulic fracturing

SUMMARY Microseismic monitoring is an important technique to obtain detailed knowledge of in-situ fracture size and orientation during stimulation to maximize fluid flow throughout the rock volume and optimize production. Furthermore, considering that the frequency of earthquake magnitudes empirically follows a power law (i.e. Gutenberg–Richter), the accuracy of microseismic event magnitude distributions is potentially crucial for seismic risk management. In this study, we analyse microseismicity observed during four hydraulic fracture treatments of the legacy Cotton Valley experiment in 1997 at the Carthage gas field of East Texas, where fractures were activated at the base of the sand-shale Upper Cotton Valley formation. We perform waveform cross-correlation to detect similar event clusters, measure relative amplitude from aligned waveform pairs with a principal component analysis, then measure precise relative magnitudes. The new magnitudes significantly reduce the deviations between magnitude differences and relative amplitudes of event pairs. This subsequently reduces the magnitude differences between clusters located at different depths. Reduction in magnitude differences between clusters suggests that some attenuation-related biases could be effectively mitigated with relative magnitude measurements. The maximum likelihood method is applied to understand the magnitude frequency distributions and quantify the seismogenic index of the clusters. Statistical analyses with new magnitudes suggest that fractures that are more favourably oriented for shear failure have lower b-value and higher seismogenic index, suggesting higher potential for relatively larger earthquakes, rather than fractures subparallel to maximum horizontal principal stress orientation.

58 GEOSCIENCES

Detecting anomalous SRF cavity behavior with unsupervised learning

We present an unsupervised learning framework for detecting anomalous superconducting radio-frequency (SRF) cavity behavior at the Continuous Electron Beam Accelerator Facility (CEBAF), emphasizing its initial performance and effectiveness. Key to the system’s success was the development of data acquisition systems (DAQs) that capture fast-sampled, information-rich signals, essential for detecting transient effects. The approach involves creating daily cavity-specific models using principal component analysis to handle variations in rf signal behavior and mitigate performance degradation from data drift. This unsupervised method eliminates the need for expensive labeling by continuously updating models with recent data. Deployed and operational for 3 months before a scheduled shutdown, the system successfully identified several issues with DAQ signals, confirming its effectiveness. Despite access to only a fraction of CEBAF’s SRF cavity signals, the framework efficiently detected several instances requiring intervention, demonstrating a significant improvement over traditional, labor-intensive methods of manual plot inspection. Published by the American Physical Society 2025

43 PARTICLE ACCELERATORS

Machine learning BPS spectra and the gap conjecture

We explore statistical properties of Bogomol’nyi-Prasad-Sommerfield q-series for strongly coupled supersymmetric theories that correspond to a particular family of three-manifolds. We discover that gaps between exponents in the -series are statistically more significant at the beginning of the -series compared to gaps that appear in higher powers of. Our observations are obtained by calculating saliencies of -series features used as input data for principal component analysis, which is a standard example of an explainable machine learning technique that allows for a direct calculation and a better analysis of feature saliencies.

97 MATHEMATICS AND COMPUTING

Machine learning inversion from scattering for mechanically driven polymers

A machine learning inversion method is developed for analyzing scattering functions of mechanically driven polymers and extracting the corresponding feature parameters, which include energy parameters and conformation variables. The polymer is modeled as a chain of fixed-length bonds constrained by bending energy, and it is subject to external forces such as stretching and shear. We generate a data set consisting of random combinations of energy parameters, including bending modulus, stretching and shear force, along with Monte Carlo-calculated scattering functions and conformation variables such as end-to-end distance, radius of gyration and off-diagonal component of the gyration tensor. The effects of the energy parameters on the polymer are captured by the scattering function, and principal component analysis ensures the feasibility of the machine learning inversion. Finally, we train a Gaussian process regressor using part of the data set as a training set and validate the trained regressor for inversion using the rest of the data. The regressor successfully extracts the feature parameters.

Gaussian process regressors

SO(3)-invariant PCA with application to molecular data

Principal component analysis (PCA) is a fundamental technique for dimensionality reduction and denoising; however, its application to three-dimensional data with arbitrary orientations -- common in structural biology -- presents significant challenges. A naive approach requires augmenting the dataset with many rotated copies of each sample, incurring prohibitive computational costs. In this paper, we extend PCA to 3D volumetric datasets with unknown orientations by developing an efficient and principled framework for SO(3)-invariant PCA that implicitly accounts for all rotations without explicit data augmentation. By exploiting underlying algebraic structure, we demonstrate that the computation involves only the square root of the total number of covariance entries, resulting in a substantial reduction in complexity. We validate the method on real-world molecular datasets, demonstrating its effectiveness and opening up new possibilities for large-scale, high-dimensional reconstruction problems.

Fraiman, Michael [Tel Aviv Univ., Tel Aviv (Israel

Impulsive Magnetic Anomaly Detection At the 100-m Scale With an Array of Induction Coil Magnetometers

We demonstrate magnetic anomaly detection (MAD) using an array of 24 commercial induction coil magnetometers with stand-off distances from a pulsed 99.8(3) kA·m 2 magnetic dipole source of 260–1200 m. The sparse array is used to estimate the magnetic dipole location, magnitude, and orientation. We demonstrate how independent component analysis (ICA) improves the accuracy and precision of the magnetometer array when estimating the dipole parameters. Using sensor responses recorded from individual source pulses, we estimate the dipole location to within 29 ± 2 m, the magnitude to within 3 ± 3 kA·m 2 , and dipole orientation error to within 19 ± 0.6°.

47 OTHER INSTRUMENTATION

Pan-Pacific low-frequency modes of sea level and climate variability

Tide gauges provide a long observational record that can inform the nature of satellite-era basin-scale sea level trends. However, common signals must be extracted from geographically sparse records. Here, by applying low-frequency component analysis (LFCA) to tide gauge records and surface climate reconstructions, we isolate three coherent modes of Pacific Ocean variability that we ascribe to: a secular, greenhouse gas–driven climate change (LFC1); a nonlinear mode of variability with a reversal around 1980, potentially linked to aerosols (LFC2); and the Interdecadal Pacific Oscillation (LFC3). Although sea level trend patterns reflect the superimposed contribution of all modes, satellite-era trends are dominated by an increasing phase of LFC2: They are thus potentially unrepresentative of both longer-term historical patterns and those expected in the future.

Science & Technology - Other Topics

VAIM-CFF: a variational autoencoder inverse mapper solution to Compton form factor extraction from deeply virtual exclusive reactions

We develop a new methodology for extracting Compton form factors (CFFs) from deeply virtual exclusive reactions such as the unpolarized DVCS cross section using a specialized inverse problem solver, a variational autoencoder inverse mapper (VAIM). The VAIM-CFF framework not only allows us access to a fitted solution set possibly containing multiple solutions in the extraction of all 8 CFFs from a single cross section measurement, but also accesses the lost information contained in the forward mapping from CFFs to cross section. We investigate various assumptions and their effects on the predicted CFFs such as cross section organization, number of extracted CFFs, use of uncertainty quantification technique, and inclusion of prior physics information. We then use dimensionality reduction techniques such as principal component analysis to visualize the missing physics information tracked in the latent space of the VAIM framework. Through re-framing the extraction of CFFs as an inverse problem, we gain access to fundamental properties of the problem not comprehensible in standard fitting methodologies: exploring the limits of the information encoded in deeply virtual exclusive experiments.

Accelerator Physics

Visual Instance-aware Prompt Tuning

Visual Prompt Tuning (VPT) has emerged as a parameter-efficient fine-tuning paradigm for vision transformers, with conventional approaches utilizing dataset-level prompts that remain the same across all input instances. We observe that this strategy results in sub-optimal performance due to high variance in downstream datasets. To address this challenge, we propose Visual Instance-aware Prompt Tuning (ViaPT), which generates instance-aware prompts based on each individual input and fuses them with dataset-level prompts, leveraging Principal Component Analysis (PCA) to retain important prompting information. Moreover, we reveal that VPT-Deep and VPT-Shallow represent two corner cases based on a conceptual understanding, in which they fail to effectively capture instance-specific information, while random dimension reduction on prompts only yields performance between the two extremes. Instead, ViaPT overcomes these limitations by balancing dataset-level and instance-level knowledge, while reducing the amount of learnable parameters compared to VPT-Deep. Extensive experiments across 34 diverse datasets demonstrate that our method consistently outperforms state-of-the-art baselines, establishing a new paradigm for analyzing and optimizing visual prompts for vision transformers.

Xiao, Xi [ORNL] (ORCID:0009000009316982)

GP-BayesOpInf

SAND2025-01851O GP-BayesOpInf is a software tool that uses algorithms to combine Gaussian process regression, principal component analysis, and linear Bayesian inference to produce a probabilistic reduced-order model for time-dependent systems. Numerical examples include the compressible Euler equations for an ideal gas, a heat diffusion process with a nonlinear reaction term, and a set of ordinary differential equations describing a compartmental model in epidemiology. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

SciDAC

PCA -BASED COMPRESSOR HARDWARE DESIGN GENERATOR IN CHISEL

SF-25-077 This repository includes a hardware generator written in Chisel that generates lossy compression hardware designs based on principal component analysis (PCA), along with a flexible testbench. It also includes a Python tool for evaluating the accuracy loss of integer quantization.

Kazutomo, Yoshi [Argonne National Laboratory (ANL)

Forced Component Estimation Statistical Method Intercomparison Project (ForceSMIP)

Anthropogenic climate change is unfolding rapidly, yet its regional manifestation can be obscured by internal variability. A primary goal of climate science is to identify the externally forced climate response from among the noise of internal variability. Separating the forced response from internal variability can be addressed in climate models by using a large ensemble to average over different possible realizations of internal variability. However, with only one realization of the real world, it is a major challenge to isolate the forced response directly in observations. In the Forced Component Estimation Statistical Method Intercomparison Project (ForceSMIP), contributors used existing and newly developed statistical and machine learning methods to estimate the forced response over 1950–2022 within individual realizations of the climate system. Participants used neural networks, linear inverse models, fingerprinting methods, and low-frequency component analysis, among other approaches. These methods were trained using large ensembles from multiple climate models and then applied to observations. Here, we evaluate method performance within large ensembles and investigate the estimates of the forced response in observations. Our results show that many different types of methods are skillful for estimating the forced response in climate models, though the relative skill of individual methods varies depending on the variable and evaluation metric. Methods with comparable skill in models can give a wide range of estimates of the forced response pattern in observations, illustrating the epistemic uncertainty in forced response estimates. ForceSMIP gives new insights into the forced response in observations, its uncertainty, and methods for its estimation.

Climate attribution

Mechanical and Thermal Forcing for Upslope Flows and Cumulus Convection over the Sierras de Córdoba

Abstract The upslope flow processes affecting the vertical extent of orographic cumulus convection are examined using observations from the Cloud, Aerosol, and Complex Terrain Interactions (CACTI) field campaign. Specifically, clear air returns from the U.S. Department of Energy (DOE) second-generation C-band scanning Atmospheric Radiation Measurement (ARM) precipitation radar (CSAPR2) are used to characterize the structure and variability of the ridge-normal (i.e., up/downslope) flow components, which transport mass to the crest of Argentina’s Sierras de Córdoba and contribute to convective initiation. Data are compiled for the entire CACTI period (October–April), including days with clear skies, shallow cumuli, cumulus congestus, and deep convection. To examine shared variability among >70 000 radar scans, we use (i) a principal component analysis (PCA) to isolate modes of variability in the upslope flow and (ii) composite analysis based on convective outcomes, determined from GOES-16 satellite observations. These data are contextualized with observed surface sensible heat fluxes, thermodynamic profiles, and synoptic-scale analysis. Results indicate distinct thermally and mechanically forced upslope flow modes, modulated by diurnal heating and synoptic-scale variations, respectively. In some instances, there is a superposition of thermal and mechanical forcing, yielding either deeper or shallower upslope flow. The composite analyses based on satellite data show that successively deeper convective outcomes are associated with successively deeper upslope flow layers that more readily transport mass to the ridge crest in conjunction with lower lifting condensation levels, facilitating convective initiation. These results help to isolate the forcing mechanisms for orographic convection and thus provide a foundation for parameterizing orographic convective processes in coarse resolution models.

Meteorology & Atmospheric Sciences

Structural features of xylan dictate reactivity and functionalization potential for bio-based materials

Plant-based materials have the potential to replace some petroleum-based products, offering compostability and biodegradability as critical advantages. Xylan-rich biomass sources are gaining recognition due to their abundance and underutilization in current industrial applications. Research of potential xylan applications has been complicated by the complex and heterogeneous structure that varies for different xylan feedstocks. Acylation is a broadly used reaction in functionalization of polysaccharides at an industrial scale. However, the efficiency of this reaction varies with the xylan source. To optimize xylan valorization, a systematic understanding of structure–reactivity relationships is essential. This study explores, characterizes, and compares various xylan feedstocks in the acylation process. Xylan feedstocks were analyzed for their chemical composition, degree of polymerization, branching, solubility, and presence of impurities. These features were correlated with xylan glycotypes’ reactivity toward functionalization with succinic anhydride in an optimized DMSO/KOH condition, achieving carboxyl contents of up to 1.46. We used principal component analysis and hierarchical clustering to identify key structural features of xylan that promote its reactivity. Our findings reveal that xylans with higher xylose content and lower degrees of branching exhibit enhanced reactivity, achieving higher carboxyl content and yields. Structural analyses confirmed successful modification, and light scattering analyses showed dramatic changes in the solution properties. Succinylation improves the solubility and film-forming properties of native xylans. This study shows key structure–reactivity relationships in xylan succinylation, establishing that low branching, high xylose content, and reduced lignin impurity enhance chemical functionalization. The results offer a framework for selecting optimal biomass feedstocks and support future efforts in genetic and synthetic biology to design plants with tunable xylan architectures. These findings advance the hemicellulose valorization for applications in coatings and packaging.

Acylation

Effects of 9.5 years warming on SOC concentration and composition in bulk soil and density fractions

Original data of whole-soil warming experiment after 9.5 years at Blodgett Forest Research Station. The Blodgett Forest is a mixed coniferous temperate forest with Mediterranean climate. The annual air temperature is 12.5℃ and the annual precipitation is 1774 mm yr-1- The soil is mesic ultic Alfisol of granitic origin, equivalent to Dystric Cambisol according to The World Reference Base for Soil Resources (WRB) system. The soil is warmed down to 1 m at + 4℃ by vertically installed heating cables. At the time of soil sampling on 1 May 2023, the whole-soil warming experiment had been running for approximately 9.5 years, from January 2014 to May 2023. The dataset includes: - Bulk_EA: C, N content, δ13C, and CN ratio of bulk soil; - Density_fractionation: organic carbon concentration, δ13C, and C/N ratio of free light fraction (fLF), occluded ligh fraction (oLF), and heavy fraction (HF); - PCA_DRIFT_AUC: original data of area under the curve (AUC) values of eight carbon bond types integrated on diffuse reflectance infrared fourier transform spectroscopy for each soil sample and soil fraction, which are consequently used for principal component analysis (PCA); - DRIFTS_stability_index: the calculation of aliphatic C–H (3000–2800 cm-1) to aromatic C=C (1670–1600 cm-1) ratios for each bulk soil sample and soil fraction. All data are provided in CSV format and can be viewed using Microsoft Excel.

Climate change