Search NASA⌕ Search

SEARCH · Search NASA

Results for “multivariate analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Insights into Tetravalent Np Speciation in HNO 3 through Spectroelectrochemistry and Multivariate Analysis

In situ optical spectroscopy, spectropotentiometry, and multivariate analysis were applied to the Np(IV) nitrate system to better understand speciation and quantify HNO 3 concentration. Thin-layer spectropotentiometry, or spectroelectrochemistry, was leveraged to isolate and stabilize Np(IV) without compromising the solution conditions and generate representative Vis-NIR absorption spectra from 0.5 to 10 M HNO 3 and benchmark the corresponding Np(IV) molar absorptivity coefficients. Spectra were described with principal component analysis (PCA) to identify the purest Np(IV) absorbance spectra among other oxidation states [e.g., Np(V/VI)] at each acid concentration and then to identify the primary sources of variance within each Np(IV) spectrum with respect to Np(IV) nitrate complexes. Then, partial least-squares regression (PLSR) and support vector regression (SVR) models were built to predict HNO 3 concentration from the Np(IV) spectral data. The nonlinear SVR model outperformed the linear PLSR model for the HNO 3 concentration predictions. Finally, the inclusion of spectra collected in edge and center point HNO 3 concentrations in the calibration set was determined to be crucial for producing models with strong predictive capabilities. The multivariate approach used in this study makes it possible to quantify HNO 3 concentration solely based on Np(IV) absorption spectra, which is essential to quantifying processing streams in various online monitoring applications.

38 RADIATION CHEMISTRY, RADIOCHEMISTRY, AND NUCLEA↗

Combining ToF‐SIMS and Multivariate Analysis to Resolve Active Sites on Ni‐Based HER Catalysts

Unambiguous identification of active sites in heterogeneous catalysis remains a major challenge, particularly for materials with ultrathin, chemically mixed surface layers. Here, we demonstrate a generalizable approach that combines time-of-flight secondary ion mass spectrometry (ToF-SIMS) with multivariate statistical analysis (principal component analysis [PCA] and multivariate curve resolution [MCR]) to resolve catalytically relevant motifs at the nanoscale. Using Ni electrodes as a model system, PCA distinguished hydroxide-enriched domains from oxide- and metal-rich regions, while MCR decomposed depth profiles and 3D images into hydroxide, oxide, and metallic layers with nanometer resolution. A unique secondary-ion fragment, NiO 3 H 3 − (m/z 108.94), emerged as a marker of hydroxide-rich environments and correlated with hydrogen evolution reaction (HER) activity across a series of Ni electrodes. Complementary density functional theory (DFT) calculations revealed that Ni(OH) 2 clusters adjacent to metallic Ni offer the most favorable water dissociation energetics, establishing the structural origin of the marker. While illustrated here for Ni-based HER, this workflow provides a broadly applicable framework to isolate and rank near-surface patterns that govern catalytic activity, thereby extending ToF-SIMS from a qualitative probe to a predictive tool for active site identification.

HER active sites↗

Evaluation of COTS Electronics by Power Spectrum Analysis and Multivariate Data Analysis

Power spectrum analysis (PSA) is a fast, non-destructive, sensitive method for examining commercial off-the-shelf ( COTS ) electronic components. These features make PSA attractive for both component screening and surveillance in support of component reliability efforts. Current analysis methods limit the utility of PSA due to the need to manually examine the results of analysis to identify anomalous parts. This study demonstrates the development and application of a workflow to automate the screening of COTS electronic components. Further, this study demonstrates the use of multivariate algorithms to assess aging of Zener diodes. These workflows can be readily extended to other components, combining the benefits of PSA and multivariate analysis to screen and evaluate COTS electronic components.

42 ENGINEERING↗

Bioinspired oxidation of benzyl alcohol: The role of environment and nuclearity of the catalyst evaluated by multivariate analysis

Inspired by copper-containing enzymes such as galactose oxidase and catechol oxidase, in which distinct coordination environments and nuclearities lead to specific catalytic activities, we summarize here the catalytic properties of dinuclear and mononuclear copper species towards benzyl alcohol oxidation using a multivariate statistical approach. Here, the new dinuclear [Cu 2 (μ-L 1 )(μ-pz)] 2+ (1) is compared against the mononuclear [CuL 2 Cl] (2), where (L 1 ) - and (L 2 ) - are the respective deprotonated forms of 2,6-bis((bis(pyridin-2-ylmethyl)amino)methyl)-4-methylphenol, and 3-((bis(pyridin-2-ylmethyl)amino)methyl)-2-hydroxy-5-methylbenzaldehyde and (pz) - is a pyrazolato bridge. Copper(II) perchlorate (CP) is used as control. The catalytic oxidation of benzyl alcohol is pursued, aiming to assess the role of the ligand environment and nuclearity. The multivariate statistical approach allows for the search of optimal catalytic conditions, considering variables such as catalyst load, hydrogen peroxide load, and time. Species 1, 2 and CP promoted selective production of benzaldehyde at different yields, with only negligible amounts of benzoic acid. Under normalized conditions, 2 showed superior catalytic activity. This species is 3.5-fold more active than the monometallic control CP, and points out to the need for an efficient ligand framework. Species 2 is 6-fold more active than the dinuclear 1, and indicates the favored nuclearity for the conversion of alcohols into aldehydes.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Multivariate analysis: An essential for studying complex glasses

Understanding the impact of individual compositional components on the devitrification of complex multicomponent glasses, for example, 10–50+ oxides, typically requires numerous studies to examine each component's impact. Here we apply exploratory data analysis (EDA) to a heterogeneous data set of silicate glasses to determine the cations’ individual and interacting effects on the crystallization of nepheline (nominally NaAlSiO 4 ). Our data consisted of 795 simulated high-level nuclear waste glasses composed of, on average, 50 oxide components. We determine the interactions in the heterogeneous data that cause deviations from the behavior found in simplified composition studies. Using both univariate and bivariate EDA techniques, we demonstrate the importance of including calculated structural glass parameters on nepheline's devitrification, including field strength, cation-to-anion radius ratio, and single-bond strength. Here, we also show that studies with simplified glass compositions may fall short in generating knowledge directly transferrable to complex glass compositions. The method used in this study has the potential to inform experimental design for simplified compositions (~6+ oxides) that can generate knowledge directly transferrable to complex, multivariable compositions. The observations reported here have broad implications for any study attempting to map the physical properties of a complex glass containing numerous cations.

36 MATERIALS SCIENCE↗

Multivariate Analysis as a Tool for Validating Tester Matching

A method of applying Principal Component Analysis, Soft Independent Modeling of Class Analysis, and statistical analysis is described that can be applied to many types of testers to ascertain how well matched the performance of the testers in the analysis are to one another or how well matched a tester is to itself at a later time. This method is most useful for situations for which the same units have not been run across the testers being analyzed for matched performance.

Multari, Rosalie A [Sandia National Laboratories (↗

Surface-Enhanced Raman Spectroscopy Combined with Multivariate Analysis for Fingerprinting Clinically Similar Fibromyalgia and Long COVID Syndromes

Fibromyalgia (FM) is a chronic central sensitivity syndrome characterized by augmented pain processing at diffuse body sites and presents as a multimorbid clinical condition. Long COVID (LC) is a heterogenous clinical syndrome that affects 10–20% of individuals following COVID-19 infection. FM and LC share similarities with regard to the pain and other clinical symptoms experienced, thereby posing a challenge for accurate diagnosis. This research explores the feasibility of using surface-enhanced Raman spectroscopy (SERS) combined with soft independent modelling of class analogies (SIMCAs) to develop classification models differentiating LC and FM. Venous blood samples were collected using two supports, dried bloodspot cards (DBS, n = 48 FM and n = 46 LC) and volumetric absorptive micro-sampling tips (VAMS, n = 39 FM and n = 39 LC). A semi-permeable membrane (10 kDa) was used to extract low molecular fraction (LMF) from the blood samples, and Raman spectra were acquired using SERS with gold nanoparticles (AuNPs). Soft independent modelling of class analogy (SIMCA) models developed with spectral data of blood samples collected in VAMS tips showed superior performance with a validation performance of 100% accuracy, sensitivity, and specificity, achieving an excellent classification accuracy of 0.86 area under the curve (AUC). Amide groups, aromatic and acidic amino acids were responsible for the discrimination patterns among FM and LC syndromes, emphasizing the findings from our previous studies. Overall, our results demonstrate the ability of AuNP SERS to identify unique metabolites that can be potentially used as spectral biomarkers to differentiate FM and LC.

60 APPLIED LIFE SCIENCES↗

Dataset_for_Conserved_macromolecular_architecture_of_Poplar_secondary_cell_walls_revealed_by_ssNMR_and_atomistic_modeling

This dataset contains solid-state 13C NMR data and atomistic molecular dynamics simulation files supporting the study of nanoscale secondary cell wall architecture across 13 genetically diverse Populus trichocarpa genotypes grown under uniform greenhouse conditions in 13C-enriched CO2 atmospheres (~89% 13C enrichment).The dataset contains two collections of solid-state 13C NMR data. (1) 200 MHz data (Bruker Avance III HD, 4 mm HX probe, 10 kHz MAS): raw Bruker TopSpin experiment folders and DMFIT-exported ascii spectra for selective and non-selective 1D 13C-13C spin diffusion experiments (3000 ms mixing) used to quantify inter-polymer spatial proximities, and short-mixing (1 ms) reference spectra used for polymeric abundance quantification by spectral deconvolution. (2) 600 MHz data (Bruker Avance III, 1.6 mm PhoenixNMR HXY probe, 30 kHz MAS): raw Bruker TopSpin experiment folders containing 2D CORD, 2D CP-INADEQUATE, and 13C/1H relaxation (T1, T1rho) experiments for all 13 genotypes, with processed Excel workbooks per experiment type. Molecular dynamics simulation code, coordinate files, and analysis scripts (NAMD/CHARMM/Python) for six atomistic cell wall models are included. Summarized ssNMR data are compiled into a single excel file and subjected to statistical analysis. Multivariate analysis code (PCA, Pearson correlation) and summary data are provided as excel worksheets and Jupyter notebooks (Python 3).

09 BIOMASS FUELS↗

Intrinsic Kinetics of Polyethylene Terephthalate Pyrolysis via Micropyrolysis and Multivariate Chromatographic Analysis

This study provides an in-depth investigation of the primary decomposition of polyethylene terephthalate (PET) via pyrolysis, employing an experimental-analytic workflow that integrates design of experiments (DoE), micropyrolysis coupled with comprehensive two-dimensional gas chromatography (GC×GC), and multivariate data analysis to verify intrinsic kinetic conditions and elucidate evolving product distributions for mapping key reaction pathways. Peaks that could not be identified using commercial spectral libraries were assigned using Mass Frontier simulations, enabling the identification of divinyl terephthalate, ethyl vinyl terephthalate, and 2-(benzoyloxy)ethyl vinyl terephthalate. A polar×polar (non-orthogonal) column set tailored for the detection of carboxylic acids enhanced the quantification of benzoic acid, 4-vinylbenzoic acid, 4-ethylbenzoic acid, and methylbenzoic acid by up to 6-fold relative to an orthogonal column combination (non-polar×mid-polar). Moreover, pyrolysis variables were systematically evaluated using a Box- Behnken design (BBD), encompassing pyrolysis temperature (500−600 °C), sample weight (50−150 μg), and carrier gas flow rate (100−300 mL min −1 ). Among these, pyrolysis temperature was the only statistically significant factor influencing product yields, ranging from 58.78 to 84.26 wt %. In contrast, neither the sample weight nor the carrier gas flow rate had a significant effect on product yields within the evaluated experimental space. At 600 °C, the major pyrolysis products were benzoic acid (up to 20.20 ± 1.46 wt %) and CO 2 (up to 21.28 ± 1.46 wt %), which can be produced through decarboxylation reactions. These findings underscore the critical importance of selecting appropriate analytical columns for the accurate quantification of heteroatomcontaining products such as carboxylic acids, which may otherwise be underestimated or undetected due to their reactivity with the stationary phase of non-polar and mid-polar columns, as well as other GC components. They also highlight the importance of selecting pyrolysis conditions for investigating the primary decomposition of PET under an isothermal kinetically limited regime.

aromatic compounds↗

Upscaling Soil Organic Carbon Measurements at the Continental Scale Using Multivariate Clustering Analysis and Machine Learning

Abstract Estimates of soil organic carbon (SOC) stocks are essential for many environmental applications. However, significant inconsistencies exist in SOC stock estimates for the U.S. across current SOC maps. We propose a framework that combines unsupervised multivariate geographic clustering (MGC) and supervised Random Forests regression, improving SOC maps by capturing heterogeneous relationships with SOC drivers. We first used MGC to divide the U.S. into 20 SOC regions based on the similarity of covariates (soil biogeochemical, bioclimatic, biological, and physiographic variables). Subsequently, separate Random Forests models were trained for each SOC region, utilizing environmental covariates and SOC observations. Our estimated SOC stocks for the U.S. (52.6 ± 3.2 Pg for 0–30 cm and 108.3 ± 8.2 Pg for 0–100 cm depth) were within the range estimated by existing products like Harmonized World Soil Database, HWSD (46.7 Pg for 0–30 cm and 90.7 Pg for 0–100 cm depth) and SoilGrids 2.0 (45.7 Pg for 0–30 cm and 133.0 Pg for 0–100 cm depth). However, independent validation with soil profile data from the National Ecological Observatory Network showed that our approach ( R 2 = 0.51) outperformed the estimates obtained from Harmonized World Soil Database ( R 2 = 0.23) and SoilGrids 2.0 ( R 2 = 0.39) for the topsoil (0–30 cm). Uncertainty analysis (e.g., low representativeness and high coefficients of variation) identified regions requiring more measurements, such as Alaska and the deserts of the U.S. Southwest. Our approach effectively captures the heterogeneous relationships between widely available predictors and the current SOC baseline across regions, offering reliable SOC estimates at 1 km resolution for benchmarking Earth system models.

58 GEOSCIENCES↗

Comparing gas composition from fast pyrolysis of live foliage measured in bench-scale and fire-scale experiments

Background: Fire models have used pyrolysis data from oxidising and non-oxidising environments for flaming combustion. In wildland fires pyrolysis, flaming and smouldering combustion typically occur in an oxidising environment (the atmosphere). Aims: Using compositional data analysis methods, determine if the composition of pyrolysis gases measured in non-oxidising and ambient (oxidising) atmospheric conditions were similar. Methods: Permanent gases and tars were measured in a fuel-rich (non-oxidising) environment in a flat flame burner (FFB). Permanent and light hydrocarbon gases were measured for the same fuels heated by a fire flame in ambient atmospheric conditions (oxidising environment). Log-ratio balances of the measured gases common to both environments (CO, CO 2 , CH 4 , H 2 , C 6 H 6 O (phenol), and other gases) were examined by principal components analysis (PCA), canonical discriminant analysis (CDA) and permutational multivariate analysis of variance (PERMANOVA). Key results: Mean composition changed between the non-oxidising and ambient atmosphere samples. PCA showed that flat flame burner (FFB) samples were tightly clustered and distinct from the ambient atmosphere samples. CDA found that the difference between environments was defined by the CO-CO 2 log-ratio balance. PERMANOVA and pairwise comparisons found FFB samples differed from the ambient atmosphere samples which did not differ from each other. Conclusion: Relative composition of these pyrolysis gases differed between the oxidising and non-oxidising environments. This comparison was one of the first comparisons made between bench-scale and field scale pyrolysis measurements using compositional data analysis. Implications: These results indicate the need for more fundamental research on the early time-dependent pyrolysis of vegetation in the presence of oxygen.

54 ENVIRONMENTAL SCIENCES↗

Novel principal component analysis tool based on python for analysis of complex spectra of time-of-flight secondary ion mass spectrometry

Time-of-flight secondary ion mass spectrometry (ToF-SIMS) is a powerful surface analysis tool, which can simultaneously provide elemental, isotopic, and molecular information with part per million (ppm) sensitivity. However, each spectrum may be composed of hundreds of ion signals, which makes the spectra data complex. Principal component analysis (PCA) is a multivariate analysis technique that has been widely used to figure out the variances among samples in ToF-SIMS spectra data analysis and is showing great success in the explanation of complex ToF-SIMS spectra. So far, several software tools have been developed for PCA of ToF-SIMS spectra; however, none of them are freely available. Such a situation leads to some difficulties in extending applications of PCA to various research fields. More importantly, it has long been challenging for common researchers to understand PCA plots and extract chemical differences among samples. In this work, we developed a new and flexible software tool (named “advanced spectra pca toolbox”) based on python for PCA of complex ToF-SIMS spectra along with an easy-to-read manual. It can generate data analysis reports automatically to explain chemical differences among samples, allowing less experienced researchers to easily understand tricky PCA results. Moreover, it is expandable and compatible with artificial intelligence/machine learning functions. Pure goethite and different lignin adsorbed goethite samples were used as a model system to demonstrate our new software tool, proving that our software tool can be readily used in complex spectra data processing. Our new software tool is open-source, convenient, flexible, and expandable. We expect this open-source tool will benefit the ToF-SIMS community.

47 OTHER INSTRUMENTATION↗

Towards the extraction of the crystal cell parameters from pair distribution function profiles

The approach based on atomic pair distribution function (PDF) has revolutionized structural investigations by X-ray/electron diffraction of nano or quasi-amorphous materials, opening up the possibility of exploring short-range order. However, the ab initio crystal structural solution by the PDF is far from being achieved due to the difficulty in determining the crystallographic properties of the unit cell. A method for estimating the crystal cell parameters directly from a PDF profile is presented, which is composed of two steps: first, the type of crystal cell is inferred using machine-learning approaches applied to the PDF profile; second, the crystal cell parameters are extracted by means of multivariate analysis combined with vector superposition techniques. The procedure has been validated on a large number of PDF profiles calculated from known crystal structures and on a small number of measured PDF profiles. The lattice determination step has been benchmarked by a comprehensive exploration of different classifiers and different input data. The highest performance is obtained using the k -nearest neighbours classifier applied to whole PDF profiles. Descriptors calculated from the PDF profiles by recurrence quantitative analysis produce results that can be interpreted in terms of PDF properties, and the significance of each descriptor in determining the prediction is evaluated. The cell parameter extraction step depends on the cell metric rather than its type. Monometric, dimetric and trimetric cells have top-1 estimates that are correct 40, 20 and 5% of the time, respectively. Promising results were obtained when analysing real nanocrystals, where unit cells close to the true ones are found within the top-1 ranked solution in the case of monometric cells and within the top-6 ranked solutions in the case of dimetric cells, even in the presence of a crystalline impurity with a weight fraction up to 40%.

36 MATERIALS SCIENCE↗

Bark morphological and chemical features are differentially correlated with disease resistance and yield in hybrid poplar taxa

In the southeastern United States, the establishment of short-rotation intensively cultured plantations of hybrid poplar has been hindered by its susceptibility to stem cankers. We evaluated the tradeoffs between biomass yield and disease tolerance in hybrid poplar genotypes belonging to P. deltoides × P. maximowiczii (DM), P. deltoides × P. nigra (DN), P. trichocarpa × P. maximowiczii (TM), and P. deltoides × P. deltoides (DD) taxa. We hypothesized that canker resistant genotypes will have thicker bark but bark thickness and biomass yield will be negatively correlated. After two growing seasons, the DD genotypes developed thicker bark compared to the genotypes of other taxa and bark thickness was not correlated with biomass yield in the DD genotypes (R 2 = 0.002). However, in the TM, DM, and DN genotypes, bark thickness was negatively correlated with biomass yield (R 2 = 0.33–0.77). Disease incidence studies revealed that the DM genotypes were most susceptible to canker whereas no disease was detected in DD genotypes. Furthermore, bark analysis conducted by Fourier transform infrared spectroscopy coupled with multivariate analysis showed that that DD genotypes to be chemically separate from the three hybrid genotypes and that bark chemistry was correlated with canker disease incidence. Taken together, these results reveal that it is possible to generate hybrid poplar genotypes with thicker bark, disease resistance, and higher biomass yields. This insight should guide further efforts to develop genetically improved hybrid poplar genotypes, both in terms of biomass yield and disease tolerance, for cultivation in the southeastern United States. Hybrid poplar cultivation in southeastern United States is hindered by its susceptibility to stem cankers. We evaluated tradeoffs between yield and canker disease resistance in various hybrid poplar genotypes. After two growing seasons, the DD genotypes showed disease resistance and developed thicker bark that was chemically distinct from the other genotypes. Bark thickness was not correlated with yield in the DD genotypes but was negatively correlated with yield in the other genotypes. These results will guide the development of hybrid poplar genotypes that are both disease resistant and high yielding for cultivation in the southeastern United States.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗