Search NASA⌕ Search

SEARCH · Search NASA

Results for “multivariate”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Electrochemical Nutrient Recovery for the Food–Energy–Water Nexus at Municipal Wastewater Facilities: Multivariate Analyses of Seasonal Sampling and Reactor Performance

Digester-equipped municipal wastewater facilities generate recycle streams with high nutrient loads that increase energy consumption and can cause environmental pollution. The reduction of these loads through electrochemical nutrient recovery (ENR) could enhance the food–energy–water nexus by producing fertilizer (struvite). This study investigated the recovery process through a 1 year sampling of recycle streams and the implementation of nutrient recovery. Time series analyses showed that P (as orthophosphate) concentration was time-variant in digester effluent streams, while N (as ammonia) concentration was time-variant in only the aerobic system. Furthermore, these two nutrient concentrations did not correlate in any of the recycle streams. Subsequent multivariate screening analyses identified anode type, NH 4 + concentration, cathodic potential, P concentration, and temperature as most significant for ENR. Finally, the optimum conditions of cathodic potential, anode area-to-volume ratio, and temperature applied to a real recycle stream resulted in 95% P recovery with 0.03 kWh/kg P. This energy consumption is significantly lower than process energy for conventional P fertilizers (1.1 kWh/kg P) and chemical recovery processes at scale (1.7–12.9 kWh/kg P). Overall, this study recommended specific process controls for nutrient recovery, expanded the variables evaluated for ENR, and demonstrated the ability to significantly impact energy demand associated with P-based fertilizers.

36 MATERIALS SCIENCE↗

A multivariate library of zirconia metal–organic frameworks with dissolved p -nitroaniline dipoles and concentration-dependent optical and dielectric response

In this work, we show how the combination of soluble non-polar and polar links allows for the preparation of multivariate metal–organic frameworks (MTV MOFs) that exhibit dipolar solid-solution behavior. We prepared a library of PIZOF-2 MOFs with varied concentrations of p -nitroaniline (PNA)-containing moiety embedded within the MTV links. This library forms a partial solid solution up to 30 mol% with input/output composition ratio of 0.536 ± 0.018. MOFs with x > 5 mol% PNA show concentration-dependent hypsochromic dipole–dipole coupling (H-coupling) in the solid-state UV-visible spectra. The MTV library also exhibits a dielectric relaxation event in the broadband dielectric spectra with activation enthalpies and entropies that vary with temperature and PNA content and follow linear Meyer–Neldel compensation relations.

Langlois, Kyle R. [Univ. of Central Florida, Orlan↗

Machine Learning Approach for Spatiotemporal Multivariate Optimization of Environmental Monitoring Sensor Locations

Abstract Long-term environmental monitoring is critical for managing the soil and groundwater at contaminated sites. Recent improvements in state-of-the-art sensor technology, communication networks, and artificial intelligence have created opportunities to modernize this monitoring activity for automated, fast, robust, and predictive monitoring. In such modernization, it is required that sensor locations be optimized to capture the spatiotemporal dynamics of all monitoring variables as well as to make it cost-effective. The legacy monitoring datasets of the target area are important to perform this optimization. In this study, we have developed a machine-learning approach to optimize sensor locations for soil and groundwater monitoring based on ensemble supervised learning and majority voting. For spatial optimization, Gaussian process regression (GPR) is used for spatial interpolation, while the majority voting is applied to accommodate the multivariate temporal dimension. Results show that the algorithms significantly outperform the random selection of the sensor locations for predictive spatiotemporal interpolation. While the method has been applied to a four-dimensional dataset (with two-dimensional space, time, and multiple contaminants), we anticipate that it can be generalizable to higher-dimensional datasets for environmental monitoring sensor location optimization.

Siddiquee, Masudur R.↗

Estimating Sparse Direct Effects in Multivariate Regression With the Spike-and-Slab LASSO

The multivariate regression interpretation of the Gaussian chain graph model simultaneously parametrizes (i) the direct effects of p predictors on q outcomes and (ii) the residual partial covariances between pairs of outcomes. We introduce a new method for fitting sparse versions of these models with spike-and-slab LASSO (SSL) priors. We develop an Expectation Conditional Maximization algorithm to obtain sparse estimates of the p × q matrix of direct effects and the q × q residual precision matrix. Our algorithm iteratively solves a sequence of penalized maximum likelihood problems with self-adaptive penalties that gradually filter out negligible regression coefficients and partial covariances. Because it adaptively penalizes individual model parameters, our method is seen to outperform fixed-penalty competitors on simulated data. We establish the posterior contraction rate for our model, buttressing our method’s excellent empirical performance with strong theoretical guarantees. Using our method, we estimated the direct effects of diet and residence type on the composition of the gut microbiome of elderly adults.

EM algorithm↗

Multiclass Classification Using Bayesian Multivariate Adaptive Regression Splines

We present a new Bayesian model for the problem of multiclass classification. In this model, the probabilities of class membership of a given observation are determined by the mean of a latent Gaussian distribution. The mean functions of this latent distribution consist of combinations of highly flexible basis functions of the inputs: multivariate adaptive regression splines (MARS), first developed for multiple regression. We use reversible jump Markov chain Monte Carlo to make inference on the classification model, including the number of basis functions. We compare the probabilistic classification performance of our proposed approach to existing methods on simulated and benchmark data, and compare uncertainty estimates on simulated data. Our proposed method compares favorably with existing Bayesian and frequentist multiclass classification methods in out-of-sample probabilistic classification, and uncertainty estimation of these probabilistic classifications. We examine the fit of the proposed method to a data set of hurricane storm surge levels near Delaware Bay, US, and conclude that sea level rise is a key contributor to damage delivered by storm surge.

97 MATHEMATICS AND COMPUTING↗

MethodOpt: a Shiny-based graphical user interface for multivariate optimization of sampling and analytical instrumentation

Method optimization is an important step in producing useful data in various experimental settings involving the use of sampling and analytical instrumentation, such as gas-chromatography mass-spectrometry or other analytical techniques. However, traditional optimization techniques often lack the sophistication of more modern optimization techniques developed in areas of applied mathematics. A graphical user interface has been developed that implements a multivariate, multi-objective optimization technique for spectra-generating sampling and analytical instrumentation, which saves substantial time and resources compared to the more traditional approaches to method development.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Transforming the $v$ World: A New Multivariate Transformer Energy Estimator for NOvA

The NOvA Transformer Energy Estimator (Transformer_EE) is a universal machine learning tool currently used to infer the incoming beam neutrino energy and the outgoing lepton energy in both near andfar detectors. It uses a unique, highly flexible framework for simultaneous multivariate prediction that supports many possible loss functions. A spectral reweighting and flattening scheme lessens training bias. A feature noising subroutine enables adversarial-like training, mitigating sensitivities to certain systematic effects at marginal resolution loss at inference time. The state of the Transformer_EE will be reviewed, and its robustness with respect to several NOvA Near and Far Detector systematics highlighted.

Tong, Leon [Minnesota U.] (ORCID:0000000231625965)↗

Data and code from: Multivariate bayesian regression model for predicting disposed ash composition at U.S. coal fired power stations

This dataset contains the code and data files needed for implementation of a Multivariate Bayesian Regression model, described in Jin et al. (2025), for the historical prediction of the chemical composition of disposed coal ash at U.S. coal fired power plants as a function of annualized coal purchase data. The integrated coal supply data file (CoalSupplyDataset.csv) represents a compilation of monthly fuel purchase records for the period 1973-2022 at major U.S. power stations. These records were obtained from the U.S. Energy Information Administration. The CSV file also contains, for each coal purchase record, the coal region of the mine as defined by the U.S. Geological Survey. Data entry errors and data gaps in the EIA records were corrected as described in Jin et al. This CSV file represents the integrated coal supply data after corrections were made. The model structure and fitting parameters are encoded in pickle file format (Bayesian.pkl). The model was developed with the coal supply data and coal ash composition data, apportioned according to the Stratified Shuffle Split for training and testing subsets. The model was built using Python and the PyMC library. Reference Publication: Jin, Z.; Huang, J.; Hower, J.C.; Hsu-Kim, H.(2025). Predictive Assessment of the Chemical Composition of Coal Ash in Reserve at U.S. Disposal Sites. Environmental Science & Technology.

Coal ash composition↗

Multivariate environmental and trait-based controls of transpiration in the Central Amazon Rainforest

Tropical forest tree mortality is increasing due to more severe droughts, yet our understanding of how tree traits and life strategies are linked to drought stress has been limited by measurement scarcity. The BIONTE (BIOmass and NuTrient Experiment) near Manaus, Brazil hosts one of the world’s largest sap flow installations, with sensors in 90 canopy trees across a wood density gradient monitored since June 2022. The 2023 El Niño drought provided a unique opportunity to evaluate how water availability impacts tree transpiration. An interpretable machine learning framework was used to study the complex interactions between transpiration and multiple environmental variables such as soil water availability and vapor pressure deficit (VPD), and how these interactions vary with wood density and individual trees. We found varying responses of transpiration from different trees during the El Niño drought. Transpiration generally increased with temperature, with stronger effects in wetter areas and in trees with low to medium wood density. However, this response was modulated by stomatal sensitivity to VPD, which constrained transpiration under high atmospheric demand, particularly in intermediate-moisture area. The inflection in transpiration rate at high temperatures (>32°C) underscores the role of stomatal and hydraulic regulation in limiting water loss and protecting trees from excessive evaporative demand. Analysis of soil water contribution to transpiration revealed unimodal patterns in wetter area, with peak contributions near 0.45 cm 3 cm -3 of surface soil water and declining or flat responses beyond that threshold, suggesting a shift from water- to energy-limited transpiration. In contrast, drier areas exhibited limited transpiration sensitivity to soil water conditions and minimal trait-based variation in VPD responses, indicating supply-limited conditions. Despite higher wood density trees being generally more resilient, this study shows diverse tree drought resilience, prompting further investigation into the specific traits and dynamics between environmental variables in regulating transpiration and other physiological processes in trees.

Drought↗

A New Simple-to-Configure Self-Perturbing Multivariable Extremum-Seeking Controller

This paper presents a new stochastic relay-based extremum-seeking controller (ESC) for multi-input-single-output (MISO) systems. The algorithm was developed with the goal of simplifying configuration to enable easier deployment to real-world problems. A solution is developed first for a static map and then adapted for a general class of dynamic systems. The number of configurable parameters is one per input channel for the static case and only one additional parameter is needed for the dynamic version. The problem of gradient identifiability is solved via the use of stochastic relay gains and a simple stability proof for the static case is presented. Simulation tests demonstrate the performance of the strategy for optimizing both static and dynamic systems.

Salsbury, Timothy [BATTELLE (PACIFIC NW LAB)]↗

Incorporation of Ion Transport Chains into Multivariate MOF for Improved Water Oxidation

The climate crisis demands clean energy technologies to cut CO 2 emissions from fossil fuels. Hydrogen fuel cells and solar-driven CO 2 reduction are promising, but both rely on efficient water oxidation. Polypyridyl ruthenium complexes are active catalysts for water oxidation; however, they exhibit poor stability and recyclability. Our group improved performance by embedding these complexes into metal−organic frameworks (MOFs). As water oxidation is pH-dependent, proton management further enhances reactivity. To address the issue, we introduced proton transfer pathways into the MOF structure. Specifically, we incorporated −SO 3 H groups onto the biphenyl linkers of UiO-67 loaded with [Ru(tpy)(dcbpy)OH 2 ]PF 6 catalyst (where tpy = 2,2′:6′,2″-terpyridine; dcbpy = 5,5-dicarboxy-2,2′- bipyridine). The sulfonated MOF exhibited a 2.5-fold increase in oxygen evolution compared to the nonsulfonated analogue. After 1 h of electrolysis, the sulfonated MOF exhibited a turnover number of 25 for oxygen evolution reaction compared to 10 for the native MOF, demonstrating the benefits of built-in proton management.

Catalysts↗

Multivariate Testing of Sampling Techniques to Address Class Imbalance in Building Use Type Classification

This study addresses the challenges inherent in building use type classification, particularly focusing on the issue of class imbalance in the training datasets for machine learning classifiers. We comprehensively analyze the efficacy of various class-balancing sampling techniques. Employing Monte Carlo simulations and Bayesian optimization, we evaluated the performance of multiple sampling methods, including Random Oversampling, Random Undersampling, SMOTE, Borderline-SMOTE, and ADASYN, across a dataset encompassing nine southeastern coastal states of the United States. Our findings reveal that simple random over- and undersampling techniques outperform more sophisticated methods. Additionally, we show inherent value in creating an imbalance in training data to effectively train a machine learning classifier for distinguishing between residential and nonresidential buildings. This study provides valuable guidance for future research on building use type classification research and lays essential groundwork for developing attribute-rich building stock datasets.

Adams, Daniel↗

Multivariate Time Series Intermittent Fault Detectionin Controller Area Network CAN

Fault detection in Controller Area Network (CAN) systems is crucial for ensuring the reliability and safety of automotive and industrial applications. This study investigates and compares the effectiveness of time series classification models for supervised fault detection in CAN data. This repository contains the code and data for our benchmarking experiment aimed at detecting intermittent faults in automotive Controller Area Network (CAN) data. The goal of this project is to compare various machine learning (ML) and deep learning (DL) models using different Time Series Cross-Validation (TSCV) techniques to evaluate their effectiveness in a streaming environment for fault detection.

Hespeler, Steven [Oak Ridge National Laboratory (O↗

Multivariate Analysis as a Tool for Validating Tester Matching

A method of applying Principal Component Analysis, Soft Independent Modeling of Class Analysis, and statistical analysis is described that can be applied to many types of testers to ascertain how well matched the performance of the testers in the analysis are to one another or how well matched a tester is to itself at a later time. This method is most useful for situations for which the same units have not been run across the testers being analyzed for matched performance.

Multari, Rosalie A [Sandia National Laboratories (↗

Surface-Enhanced Raman Spectroscopy Combined with Multivariate Analysis for Fingerprinting Clinically Similar Fibromyalgia and Long COVID Syndromes

Fibromyalgia (FM) is a chronic central sensitivity syndrome characterized by augmented pain processing at diffuse body sites and presents as a multimorbid clinical condition. Long COVID (LC) is a heterogenous clinical syndrome that affects 10–20% of individuals following COVID-19 infection. FM and LC share similarities with regard to the pain and other clinical symptoms experienced, thereby posing a challenge for accurate diagnosis. This research explores the feasibility of using surface-enhanced Raman spectroscopy (SERS) combined with soft independent modelling of class analogies (SIMCAs) to develop classification models differentiating LC and FM. Venous blood samples were collected using two supports, dried bloodspot cards (DBS, n = 48 FM and n = 46 LC) and volumetric absorptive micro-sampling tips (VAMS, n = 39 FM and n = 39 LC). A semi-permeable membrane (10 kDa) was used to extract low molecular fraction (LMF) from the blood samples, and Raman spectra were acquired using SERS with gold nanoparticles (AuNPs). Soft independent modelling of class analogy (SIMCA) models developed with spectral data of blood samples collected in VAMS tips showed superior performance with a validation performance of 100% accuracy, sensitivity, and specificity, achieving an excellent classification accuracy of 0.86 area under the curve (AUC). Amide groups, aromatic and acidic amino acids were responsible for the discrimination patterns among FM and LC syndromes, emphasizing the findings from our previous studies. Overall, our results demonstrate the ability of AuNP SERS to identify unique metabolites that can be potentially used as spectral biomarkers to differentiate FM and LC.

60 APPLIED LIFE SCIENCES↗

AI-Enabled Operations at Fermi Complex: Multivariate Time Series Prediction for Outage Prediction and Diagnosis

The Main Control Room of the Fermilab accelerator complex continuously gathers extensive time-series data from thousands of sensors monitoring the beam. However, unplanned events such as trips or voltage fluctuations often result in beam outages, causing operational downtime. This downtime not only consumes operator effort in diagnosing and addressing the issue but also leads to unnecessary energy consumption by idle machines awaiting beam restoration. The current threshold-based alarm system is reactive and faces challenges including frequent false alarms and inconsistent outage-cause labeling. To address these limitations, we propose an AI-enabled framework that leverages predictive analytics and automated labeling. Using data from $2,703$ Linac devices and $80$ operator-labeled outages, we evaluate state-of-the-art deep learning architectures, including recurrent, attention-based, and linear models, for beam outage prediction. Additionally, we assess a Random Forest-based labeling system for providing consistent, confidence-scored outage annotations. Our findings highlight the strengths and weaknesses of these architectures for beam outage prediction and identify critical gaps that must be addressed to fully harness AI for transitioning downtime handling from reactive to predictive, ultimately reducing downtime and improving decision-making in accelerator management.

Jain, Milan [PNL, Richland] (ORCID:000000021676111↗