Search NASA⌕ Search

SEARCH · Search NASA

Results for “multivariate”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

Search for dark matter production in association with bottom quarks and a lepton pair in proton-proton collisions at $\sqrt{s}=13$ TeV

A search is performed for dark matter produced in association with bottom quarks and a pair of electrons or muons in data collected with the CMS detector at the LHC, corresponding to 138 fb −1 of integrated luminosity of proton-proton collisions at a center-of-mass energy of 13 TeV. For the first time at the LHC, the associated production of a bottom quark-antiquark pair and a new heavy neutral Higgs boson (H) that subsequently decays into a leptonically decaying Z boson and a pseudoscalar (a) is explored. The latter acts as a dark matter mediator in the context of the two Higgs doublet model plus a pseudoscalar (2HDM+a). Multivariate techniques that target a wide range of mass configurations for the H and a particles are used. The observations are consistent with the expectations from standard model processes. Upper limits at 95% confidence level are set on the product of cross section and branching fraction of the new particles, ranging from 10 −2 pb for an H mass of 400 GeV to 10 −3 pb for an H mass of 2000 GeV. Constraints on the parameter space of a benchmark 2HDM+a model are derived and compared with expectations in the context of cosmological predictions.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Search for dark matter produced in association with one or two top quarks in proton-proton collisions at $\sqrt{\text{s}}$ = 13 TeV

A search is performed for dark matter (DM) produced in association with a single top quark or a pair of top quarks using the data collected with the CMS detector at the LHC from proton-proton collisions at a center-of-mass energy of 13 TeV, corresponding to 138 fb −1 of integrated luminosity. An excess of events with a large imbalance of transverse momentum is searched for across 0, 1 and 2 lepton final states. Novel multivariate techniques are used to take advantage of the differences in kinematic properties between the two DM production mechanisms. No significant deviations with respect to the standard model predictions are observed. The results are interpreted considering a simplified model in which the mediator is either a scalar or pseudoscalar particle and couples to top quarks and to DM fermions. Axion-like particles that are coupled to top quarks and DM fermions are also considered. Expected exclusion limits of 410 and 380 GeV for scalar and pseudoscalar mediator masses, respectively, are set at the 95% confidence level. A DM particle mass of 1 GeV is assumed, with mediator couplings to fermions and DM particles set to unity. A small signal-like excess is observed in data, with the largest local significance observed to be 1.9 standard deviations for the 150 GeV pseudoscalar mediator hypothesis. Because of this excess, mediator masses are only excluded below 310 (320) GeV for the scalar (pseudoscalar) mediator. The results are also translated into model-independent 95% confidence level upper limits on the visible cross section of DM production in association with top quarks, ranging from 1 pb to 0.02 pb.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Measurement of the top quark pair production cross section in PbPb collisions at $\sqrt{s_{\mathrm{NN}}}=5.36$ TeV

The inclusive cross section for top quark pair ($\mathrm{t}\overline{\mathrm{t}}$) production in lead-lead (PbPb) collisions is reported for the first time at a center-of-mass energy per nucleon pair of 5.36 TeV. The analysis uses data corresponding to an integrated luminosity of 1.58 nb −1 collected by the CMS experiment at the CERN LHC in 2023. The $\mathrm{t}\overline{\mathrm{t}}$ production cross section, ${\sigma}_{\mathrm{t}\overline{\mathrm{t}}}={3.42}_{-0.51}^{+0.54}{\left(\mathrm{stat}\right)}_{-0.43}^{+0.50}\left(\mathrm{syst}\right)$ μb, is measured in dilepton final states using a fit to a multivariate discriminator that combines the decay electron and muon kinematic properties with the multiplicity of bottom quark jets. The result is consistent with perturbative quantum chromodynamics calculations at next-to-next-to-leading order (NNLO) accuracy employing several nuclear parton distribution functions. In addition, the Drell–Yan production cross section (σ DY ) for dilepton masses above 10 GeV and the ratio of $\mathrm{t}\overline{\mathrm{t}}$ to DY cross sections $\left({R}_{\mathrm{t}\overline{\mathrm{t}}/\mathrm{DY}}\right)$ are found to be compatible with the NNLO predictions. The observables ${\sigma}_{\mathrm{t}\overline{\mathrm{t}}}$, σ DY , and ${R}_{\mathrm{t}\overline{\mathrm{t}}/\mathrm{DY}}$ are measured separately for central and semicentral PbPb collisions to investigate for the first time the dependence of top quark production on the collision impact parameter.

Heavy Ion Experiments↗

Search for the production of a Higgs boson in association with a single top quark in pp collisions at $\sqrt{s}=13$ TeV with the ATLAS detector

A search for the production of a Higgs boson in association with a single top quark, tH, is presented. The analysis uses proton-proton collision data corresponding to an integrated luminosity of 140 fb −1 at a centre-of-mass energy of 13 TeV, collected by the ATLAS detector at the LHC. The search targets Higgs-boson decays into $b\bar{b}$, WW * , ZZ * , and ττ, accompanied by an isolated lepton (electron or muon) from the top-quark decay. Multivariate techniques are employed to enhance the separation between signal and background processes. The observed signal strength, μ tH , defined as the ratio between the measured cross-section and the predicted Standard Model value, is μ tH = 8.1 ± 2.6 (stat.) ± 2.0 (syst.). The significance of the observed (expected) signal above the background-only expectation is 2.8 (0.4) standard deviations. The corresponding observed (expected) upper limit at the 95% confidence level on the tH cross-section is found to be 13.9 (6.1) times the value predicted by the Standard Model. An interpretation with an inverted sign of the top-quark Yukawa coupling is performed, and the signal strength and corresponding limit are reported.

Hadron-Hadron Scattering↗

Calculation of machine precision second order derivatives using dual-complex numbers

It is well known that both complex and dual numbers can be employed to obtain machine precision first-order derivatives; however, neither, on their own, can compute machine precision 2nd order derivatives. To address this limitation, it is demonstrated in this paper that combined dual-complex numbers can be used to compute machine precision 1st and 2nd order derivatives. The dual-complex approach is simpler than utilizing multicomplex or hyper-dual numbers as existing dual libraries can be used as is or easily augmented to accept complex numbers, and the complexity of developing, integrating, and deploying multicomplex or hyper-dual libraries is avoided. The efficacy of this approach is demonstrated for both univariate and multivariate functions. Finally, source code examples using the Python, Julia, and Mathematica languages are provided as supplemental material.

97 MATHEMATICS AND COMPUTING↗

Enhancing the Range and Reliability of the Spacer Layer Imaging Method

The spacer layer imaging method (SLIM) is widely used to measure the thickness of additive and lubricant films, in lubricant development and evaluation, and for fundamental research into elastohydrodynamic lubrication and tribofilm formation mechanisms. The film thickness measurement, as implemented on several popular tribometers, provides powerful, non-destructive in-situ mapping of film topography with nanometre-scale height sensitivity. However, the results can be highly sensitive to experimental procedure, machine condition, and image analysis, in some cases reporting unphysical film thickness trends. The prevailing image analysis techniques make it challenging to interrogate these errors, often hiding their multivariate nonlinear behaviour from the user by spatial averaging. Herein, several common ‘silent errors’ in the SLIM measurement, including colour matching to incorrect fringe orders, and colour drift due to the optical properties of the system or film itself, are discussed, with examples. A robust suite of novel a priori and a posteriori methods to address these issues, and to improve the accuracy and reliability of the measurement, are also presented, including a novel, computationally inexpensive circle-finding algorithm for automated image processing. In combination, these methods allow reliable mapping of films up to at least 800 nm in thickness, representing a significant milestone for the utility of SLIM applied to elastohydrodynamic contact.

EHL film geometry↗

Data-Driven Insights into the Structural Essence of Plasticity in High-Entropy Alloys

The heterogeneous mechanical response of a crystalline alloy with multiple principal elements was investigated using molecular dynamics simulations. The local configuration of the alloy in its quiescent state was characterized by the variables derived from the gyration tensor and the atomic electronegativity. A multivariate analysis identified the geometric and chemical factors that influenced the atomic packing variations. Further, upon straining, the non-affine displacement exhibited spatial heterogeneity. A statistical correlation was established between the local yield events and the specific features of the local configuration. Our findings, validated by the performance metrics analysis, provided a structural criterion for the instability mechanisms in high-entropy alloys (HEAs) and enhanced the understanding of their plasticity.

36 MATERIALS SCIENCE↗

Reducing systematic bias in machine learning applications to J/ψ signal extraction in high-energy nuclear physics

Machine learning techniques are increasingly used in high-energy nuclear physics because they can exploit multivariate correlations more efficiently than conventional cut-based analyses. A central challenge is the construction of training samples that faithfully reproduce the detector response observed in data. Signal samples are usually derived from detector simulations; therefore, mismatches between simulation and data can degrade classifier performance and introduce systematic biases. This work presents two practical correction procedures, namely cumulative distribution function (CDF) mapping and a shift-and-scale transformation, to align simulated signal features with those measured in data. Their performance is demonstrated with $J$/$\psi$ yield measurements in $\sqrt{s_{nn}}$ = 200 GeV Ru+Ru and Zr+Zr collisions recorded by STAR. A set of self-consistency tests shows that these procedures substantially suppress the systematic bias associated with data-simulation discrepancies in machine-learning-based signal extraction.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Reduced-order modeling for efficient cross section library development in high-temperature gas reactor pebble-bed depletion analysis

Accurate modeling of running-in and equilibrium conditions in pebble-bed reactors (PBRs) requires precise microscopic multigroup neutron cross sections. In Griffin, deterministic neutronics calculations rely on multivariate interpolation over large cross section libraries, resulting in significant memory usage and performance bottlenecks. This work, together with a companion paper on Griffin integration, explores reduced-order models (ROMs) to replace interpolation with lightweight surrogates. Several ROM techniques are benchmarked, with deep neural networks (DNNs) demonstrating superior memory efficiency, scalability, and predictive accuracy. A total of 295 DNNs were trained to build a comprehensive isotope library, integrated into Griffin through a custom LibTorch interface for depletion analysis. Initial results demonstrate that DNN-based ROMs drastically reduce memory demands while preserving accuracy, enabling finer tabulations and additional state variables without overhead. In conclusion, the framework also supports online cross section generation and real-time DNN updates through transfer learning, improving fidelity by capturing self-shielding and evolving nuclide compositions during burnup.

22 - GENERAL STUDIES OF NUCLEAR REACTORS↗

Local practically safe extremum seeking with assignable rate of attractivity to the safe set

We present Assignably Safe Extremum Seeking (ASfES), an algorithm designed to minimize a measured, static objective function while maintaining a measured, static metric of safety (a control barrier function or CBF) to be positive in a practical sense. We ensure that for trajectories with safe initial conditions, the violation of safety can be made arbitrarily small through appropriately chosen design constants. We also guarantee an assignable “attractivity” rate: from unsafe initial conditions, the trajectories approach the safe set, in the sense of the measured CBF, at a rate no slower than a user-assigned rate. Similarly, from safe initial conditions, the trajectories approach the unsafe set, in the sense of the CBF, no faster than the assigned attractivity rate. The feature of assignable attractivity is not present in the semiglobal version of safe extremum seeking, where the semiglobality of convergence is achieved by slowing the adaptation. We also demonstrate local convergence of the parameter to a neighborhood of the minimum of a quadratic objective function constrained to the safe set with a linear CBF. The ASfES algorithm and analysis are multivariable, but we also extend the algorithm to a Newton-Based ASfES scheme which we show is only useful in the scalar case. The proven properties of the designs are illustrated through simulation examples.

42 ENGINEERING↗

PFAS remediation: Evaluating the infrared spectra of complex gaseous mixtures to determine the efficacy of thermal decomposition of PFAS

Due to their widespread production and known environmental contamination, the need for the detection and remediation of per- and polyfluoroalkyl substances (PFAS) has grown quickly. While destructive thermal treatment of PFAS at low temperatures (e.g., 200 to 500oC) is of interest due to lower energy and infrastructure requirements, the range of possible degradation products remains underexplored. To better understand the low temperature decomposition of PFAS species, we have coupled gas-phase infrared spectroscopy with a multivariate curve resolution (MCR) analysis and a database of high-resolution PFAS infrared reference spectra to detect and quantify a complex mixture resulting from potassium perfluorooctanesulfonate (PFOS-K) decomposition. Nine prevalent decomposition products (namely smaller perfluorocarbon species) are identified and quantified.

54 ENVIRONMENTAL SCIENCES↗

Chemistry imaging and distribution analysis of rare earth elements in coal using LIBS and LA-ICP-MS instruments

Currently, demand for rare earth elements (REEs) increased significantly. Coal is actively evaluated as potential economic sources for extraction of REEs. Here, in this work, laser-induced breakdown spectroscopy (LIBS) was evaluated for rapid estimation of REEs content and their distribution in the natural coal samples. The results were compared with similar laser ablation–inductively coupled plasma–mass spectrometry (LA-ICP-MS) measurements. Thirteen coal samples (nine standard samples and five natural samples) were used in this study. Powder samples were pressed into pellets while coal chunks were directly ablated for data recording. Pellets of the powder standard samples were used to optimize the data acquisition system and then data recorded with this optimized system was used to identify the proper data acquisition and analysis models. After establishing the proper data acquisition system and analysis model using the standard samples, natural coal samples in powder form and their chunks were utilized to record LIBS and LA-ICP-MS spectra. Multivariate calibration models were developed using four of the natural samples, which were evaluated by predicting the REE content in the fifth sample. Principal component analysis was performed on the LIBS data obtained from the natural samples and it classified all the samples with high accuracy. Two-dimensional (2D) elemental mapping on coal chunk samples was also performed using both LIBS and LA-ICP-MS to study the distribution of REEs in the samples. The resulting elemental images and their correlations can be used to infer mineral distributions.

01 COAL, LIGNITE, AND PEAT↗

Leveraging large language models to address data scarcity in machine learning for graphene synthesis

Machine learning in experimental materials science faces significant challenges due to the scarcity of data, which are costly and time-consuming to generate, particularly when relying on in-house experiments. Literature data mining offers a potential solution but introduces issues like mixed data quality, inconsistent formats, and non-uniform reporting of synthesis parameters, resulting in partially missing and heterogeneous features across the dataset. Here, we propose data imputation and feature engineering methods that employ pre-trained large language models (LLMs) to enhance machine learning performance on scarce, heterogeneous datasets, demonstrated on graphene CVD synthesis data and the ML-HydPARK hydrogen storage dataset. GPT models perform data imputation via tailored prompting and semantic normalization of inconsistently reported features through embeddings, for example, to harmonize the complex nomenclature of CVD substrates. Beyond yielding more diverse and richer feature representations than traditional methods such as K-nearest neighbors (KNN) and Multivariate Imputation by Chained Equations (MICE), LLM-based data imputation is evaluated against dataset characteristics and prompting strategies. We vary the level of autonomy granted to the LLM, from generic prompting that leverages pre-trained knowledge for autonomous data generation to data-informed prompting that constrains outputs using target-specific information, and demonstrate which level of autonomy yields superior imputation performance across datasets and feature types. The proposed data engineering methods markedly improve downstream performance; for example, in graphene layer number classification using a support vector machine (SVM), binary accuracy increases from 39% to 65% and ternary accuracy from 52% to 72%. Fine-tuning experiments on both datasets show that combining our proposed LLM-based data imputation and feature encoding methods with numerical machine learning predictors outperforms standalone fine-tuned LLM predictors in data-scarce settings. The proposed strategies emphasize data enhancement techniques rather than refining learning architectures or regularizing loss functions, offering a broadly applicable framework for improving machine learning performance on scarce, inhomogeneous datasets.

Chemical vapor deposition↗

Mapping wall-to-wall fractional cover of Arctic tundra plant functional types in Alaska using 20-m spatial resolution satellite imagery and harmonized plot observations

Estimates of fractional cover (fCover) across given land surfaces are used to assess, and often model, vegetation composition and diversity, which are crucial for understanding the health and functioning of terrestrial ecosystems. Remote sensing provides a useful means for scaling local, plot-measured fCover estimates to regional scales. Leveraging a recently synthesized and harmonized plot database, this study generated wall-to-wall maps of fCover for six Alaskan-Arctic plant functional types (PFT), including non-vascular plants, forbs, graminoids, and deciduous and evergreen shrubs, using 20-m satellite data (Sentinel-1, Sentinel-2, ArcticDEM) using a machine learning regression approach, specifically the random forest (RF) algorithm, which is well-suited for handling nonlinear relationships and high-dimensional satellite datasets. This study additionally addressed the spatio-temporal inconsistencies e.g., sampling scale, plot size, and collection year in plot measured fCover by adopting a multivariate outlier detection approach—Cook’s distance—to identify high-quality plots for model training and validation. Our approach achieves high accuracy (R 2 = 0.59–0.93, root mean squared errors = 0.02–0.10 for all PFTs) between plot-observed and satellite-derived fCover when using high-quality plot samples. The mapped fCover characterizes the spatial patterns of different PFTs across the tundra biome at a 20-m resolution, providing key information needed for improved representation of Arctic tundra vegetation in terrestrial biosphere models to better understand climate-vegetation feedback across the Arctic tundra.

Arctic tundra↗

Feedforward-feedback ammonia control at a water resource recovery facility based on a digital twin with hybrid model

Ammonia-based aeration control (ABAC) at full-scale Water Resource Recovery Facilities (WRRFs) can be challenged by diurnal loading and transport delays. This work addressed these challenges using a hybrid feedforward–feedback controller built on Activated Sludge Model 1 (ASM1), marking the first full-scale deployment to pair a mechanistic feedforward core with data-driven corrections. The objectives were to improve ammonia setpoint tracking, assess performance of the mechanistic model when enhanced with data-driven corrections, and document full-scale operation. The hybrid model incorporates two data-driven components: (1) a Mechanistic Error Forecasting Engine (MEFE), consisting of a multivariate linear regressor and a long short-term memory (LSTM) ensemble. Defying expectations, low-parameter models outperformed more complex alternatives, reducing the mechanistic error by 71%. (2) A Residual Oscillation Forecasting Engine (ROFE), based on Fast Fourier Transform, reduced the remaining error by another 35%. Two proportional–integral (PI) feedback loops further (i) trim the feedforward output and (ii) eliminate residual controller error in the final aerobic zone. In full-scale operation, the controller reduced mean-squared error (MSE) by 94% over the baseline and produced more stable dissolved oxygen (DO) setpoints. Overall, it was proven that layering multi-timescale data-driven models on a mechanistic core can yield reliable ABAC performance at WRRFs.

54 ENVIRONMENTAL SCIENCES↗

Unraveling plant phenotype to genotype associations with daily hyperspectral traits in Populus trichocarpa

Hyperspectral remote sensing is a powerful, high-throughput phenotyping tool that quantifies physiologically and structurally relevant wavelengths across diverse genotypes and over varying temporal scales. In this study, we combined tower-based continuous hyperspectral sensing with genome-wide association studies to analyze 1423 wavebands (400-900 nm) and derivative vegetation indices across 505 genotypes and the genetic architecture of hyperspectral phenotypes over time in Populus trichocarpa Torr. & Gray grown under field conditions. Wavelengths related to chlorophyll and carotenoid absorption spectra exhibited the strongest genetic variation resulting in 98 significant SNP associations. Notably, we found substantial overlap in genetic association between the blue and red spectral regions, indicative of carotenoids and chlorophyll, respectively, and identified more than 10 candidate genes associated with chloroplast function, underpinning photosynthetic activity. Furthermore, fluctuations in associations for vegetative indices, such as the chlorophyll:carotenoid index (CCI), across the growing season reveal a temporally dynamic genetic architecture of physiological traits associated with fall senescence of this temperate tree species. Finally, we also observed correlations (spearman rho = 0.3, p < 1x10 −8 ) between individual wavebands or vegetative indices and growth rate, assessed as the relative change of tree height over the growing season. The growth rate prediction was substantially improved by a regularization multivariate model (spearman rho>0.5, p < 1x10 −16 ), reinforcing the value of hyperspectral measurements for predicting traits linked to tree productivity. These findings highlight the potential of high-throughput, rapid, hyperspectral genome wide association studies GWAS to uncover physiologically meaningful genetic variation and offer promising insights for future acceleration for plant breeding.

09 BIOMASS FUELS↗

Real-time monitoring of trace noble gases using laser-induced breakdown spectroscopy—An investigation of the impact of bulk gas on plasma properties and sensitivity

The impact of Ar and He bulk gases on laser-induced breakdown spectroscopy (LIBS) real-time monitoring of trace Xe and Kr was assessed. LIBS is being developed as a monitoring tool for measuring noble gas transport in molten salt systems, in which traditional sensors may face challenges associated with radiation, corrosive materials, and/or mixed phases. The plasma temperature and electron densities of LIBS plasmas were measured in both static and various flowing Ar and He streams (0–5 L min −1 ). The use of an Ar bulk gas resulted in higher plasma temperature, greater electron densities by an order of magnitude, and extended plasma lifetime compared with when He bulk gas was used. Gas flow rate was found to have little impact on plasma temperature; however, its effect on electron density was significant, indicating the need to consider flow rate–specific models. Matrix effects on emission peaks were reported for both bulk gases. Due to these matrix effects, multivariate models were developed for Xe and Kr ranging from 0 to 700 ppm in both bulk gases. Although the predictive behavior was similar (root mean square error of prediction ranging from 11.1 to 20.6 ppm), the limits of detection were superior in He (Xe: 22.9 ppm, Kr: 30.4 ppm). Furthermore, these models were employed in demonstrative real-time tests (>1 h), which showed strong predictive precision (relative standard deviation <5 %) regardless of the bulk gas. Ultimately, this study provides a guide for the considerations required when developing gaseous LIBS models for real-time monitoring.

Gas flow effects↗

Advanced Method Optimization for Sampling and Analysis Instrumentation

This work presents a generalized approach for analytical method optimization that branches the gap between techniques historically employed and accurate modern optimization techniques suitable for various applications. The novelty of the described strategy is the utilization of multivariate, multiobjective optimization with Karush-Kuhn-Tucker conditions to bound the optimization space to solutions within the physical limitations of instrumentation. Briefly, the basic steps outlined in this paper are to (1) determine the objective(s) that should be maximized or minimized based on the goals of the analytical application, (2) conduct a screening experiment, (3) perform ANOVA to determine the parameters which have a statistically significant effect on the objective, (4) conduct an experiment (e.g., Box-Behnken design) to collect data for fitting the objective equation, and (5) determine the physical constraints of the parameters and solve the Lagrangian to determine the optimal method parameters. A broad approach to optimization target selection allows for robust method tuning to develop improved data sets amenable for chemometrics and machine learning algorithm development. Gas chromatography-mass spectrometry was selected as a use case due to its broad use across scientific fields and time-consuming method development involving numerous parameters. In conclusion, this strategy can reduce the cost of research, improve data quality, and enable the rapid development of new analytical technique.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗