Search NASA⌕ Search

SEARCH · Search NASA

Results for “factorization machine”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 217 records · Page 12

Single Photon Searches at ICARUS with SPINE

The MiniBooNE Low-Energy Excess (LEE) of electron-like events from the Booster Neutrino Beam (BNB) has puzzled neutrino physicists for decades. One possible explanation has been an unpredicted excess of neutral current (NC) $\Delta$ resonance interactions with a subsequent radiative decay. An increase in the rate of NC $\Delta\rightarrow N\gamma$ events by a factor of 3.18 could explain the LEE seen by MiniBooNE. ICARUS also sees neutrinos from the BNB and can check the rate of these single photon events. The SPINE particle physics reconstruction suite leverages deep neural networks (DNNs) to optimize particle reconstruction and identification in Liquid Argon Time Projection Chambers (LArTPCs). I present preliminary findings on the effectiveness of these machine learnign (ML) techniques for single photon event reconstruction at ICARUS.

Hausner, Harry [Fermilab] (ORCID:0000000188932280)↗

Machine learning based prediction of airflow maldistribution in air-to-refrigerant heat exchangers

Flow maldistribution is a common challenge in heat exchanger (HX) design and particularly important for air-to-refrigerant geometries where capacity losses can approach 65%. This has a major impact on central air conditioning systems, as compact duct design motivates the use of A-type HXs which are known to be affected by airflow maldistribution. Because velocity profiles are difficult to predict, components are often oversized leading to increased material cost, system footprint, and refrigerant charge. Several studies detail airflow maldistribution for individual HXs and packages, but findings cannot always be extrapolated to new designs. In this work, a machine learning (ML) based flow profile prediction framework is developed and applied to two common package configurations: (i) A-type and (ii) U-type HXs, across a broad range of HX geometries and flow rates. Porous media CFD simulations are validated against independent data for both package types as well as comprehensive in house measurements for a finless geometry with shape optimized non-round tubes, which validates the framework for new heat transfer surfaces. The ML models are trained on the porous media CFD simulations, predicting volumetric flow rate (VFR) within 1.1% and 1.9% with maximum relative L 2 norm errors of 0.48 and 0.65, respectively, while also delivering 10 5 speed up factor compared to full porous media CFD. HX level simulations show an up to 9% reduction in heat transfer from flow maldistribution, with greater losses occurring at smaller half apex angles. This framework enables rapid and highly accurate prediction of airflow maldistribution induced capacity degradation.

42 ENGINEERING↗

Kernel methods for evolution of generalized parton distributions

Generalized parton distributions (GPDs) characterize the 3-dimensional structure of hadrons, combining information about their internal quark and gluon longitudinal momentum distributions and transverse position within the hadron. The dependence of GPDs on the factorization scale Q 2 allows one to connect hard exclusive processes involving GPDs at disparate energy and momentum scales, which is needed in global analyses of experimental data. Here, in this work, we explore how finite element methods can be used to construct fast and differentiable Q 2 evolution codes for GPDs in momentum space, which can be used in a machine learning framework. We show numerical benchmarks of the methods' accuracy, including a comparison to an existing evolution code from PARTONS/APFEL++, and provide a repository where the code can be accessed.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

wa-hls4ml and lui-gnn: A benchmark and GNN-based surrogate model for hls4ml resource and latency estimation

As machine learning (ML) increasingly serves as a tool for addressing real-time challenges in scientific applications, the development of advanced tooling has significantly reduced the time required to iterate on various designs. These advancements have solved major obstacles, but also exposed new challenges. For example, processes that were not previously considered bottlenecks, such as model synthesis, are now becoming limiting factors in the rapid iteration of designs. To reduce these emerging constraints, multiple efforts are being launched toward designing an ML-based surrogate model that estimates resource usage of synthesized accelerator architectures. This model would reduce the design iteration time, especially when designing within a set of given hardware constraints. This approach shows considerable potential, but as it stands, the effort is early and would benefit from coordination and standardization to assist future work as it emerges. We introduce wa-hls4ml, a benchmark for ML accelerator resource and latency estimation, and its corresponding initial dataset of more than 100,000 fully connected neural networks, all synthesized using hls4ml and targeting Xilinx FPGAs. In addition to the resource utilization and latency data provided, the dataset includes generated artifacts and log files for many of the synthesized neural networks, in order to support future research in ML-based code generation. The benchmark evaluates the performance of resource and latency predictors against several common ML model architectures, primarily originating from scientific domains, as exemplar models, as well as the average performance across a subset of the dataset. We measure the performance of a given predictor model through multiple metrics, including $R^2$ score and SMAPE on regression tasks, as well as inference time to further characterize the estimator under test. Additionally, we introduce the latency/utilization inference graph neural network (lui-gnn), a surrogate model that uses a graph neural network to represent input architectures in the form of a directed graph. This graph representation allows for a diverse set of model architectures to all be effectively handled by a surrogate model. We present the architecture and performance of the model, as evaluated by the new proposed benchmark, including SMAPE, $R^2$ score, and inference times, and find that lui-gnn generally predicts latency and utilization for the 75\% quantile within several percent of the synthesized resources on the synthetic test dataset, indicating that this approach of estimating resource and latency via a surrogate models has promise and warrants further research.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Sparsified time-dependent Fourier neural operators for fusion simulations

This paper presents a sparsified Fourier neural operator for coupled time-dependent partial differential equations (ST-FNO) as an efficient machine learning surrogate for fluid and particle-based fusion codes such as NIMROD (Non-Ideal Magnetohydrodynamics with Rotation - Open Discussion) and GTC (Gyrokinetic Toroidal Code). ST-FNO leverages the structures in the governing equations and utilizes neural operators to represent Green's function-like numerical operators in the corresponding numerical solvers. Once trained, ST-FNO can rapidly and accurately predict dynamics in fusion devices compared with first-principle numerical algorithms. In general, ST-FNO represents an efficient and accurate machine learning surrogate for numerical simulators for multi-variable nonlinear time-dependent partial differential equations, with the proposed architectures and loss functions. The efficacy of ST-FNO has been demonstrated using quiescent H-mode simulation data from NIMROD and kink-mode simulation data from GTC. The ST-FNO H-mode results show orders of magnitude reduction in memory and central processing unit usage in comparison with the numerical solvers in NIMROD when computing fields over a selected poloidal plane. The ST-FNO kink-mode results achieve a factor of 2 reduction in the number of parameters compared to baseline FNO models without accuracy loss.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Machine-learning-assisted deciphering of microstructural effects on ionic transport in composite materials: A case study of Li 7 La 3 Zr 2 O 12 -LiCoO 2

The effective diffusivity of ionic species in multiphase materials is critical for the design and function of composite materials for electrochemical energy storage. In practice, effective diffusivity depends sensitively not only on the intrinsic diffusivities of constituting materials but also on their topological arrangement; nevertheless, these coupled contributions are oversimplified in most analytical models. Here, we combine atomistically informed mesoscale modeling and machine learning (ML) analysis to unravel how such features affect effective diffusivity in two-phase composites. Using the Li 7 La 3 Zr 2 O 12 -LiCoO 2 composite solid-state battery cathode as a model system, we compute effective diffusivity for 600 distinct dense polycrystalline microstructures with different topological configurations of grains, grain boundaries, and heterointerfaces. We verify that in addition to atomic-scale variabilities, microstructural feature diversity can significantly impact effective transport properties. Across the ensemble of test microstructures, this often results in bimodal distributions of effective diffusivity that encompass two qualitatively distinct operating mechanisms, which we identify via flux analysis. An ML approach reveals that the most critical determining factors for effective diffusivity are the connectivity of bulk phases and their heterointerfaces. The role of ionic mobility at the heterointerfaces is also discussed. These insights highlight the combined importance of microstructure and interface engineering in tuning the transport properties of ionic species in composite materials. In conclusion, our framework can also be extended for understanding generic microstructure-property relationships in other complex multiphase materials.

25 ENERGY STORAGE↗

Direct Measurement of ICRF-Enhanced Plasma Potentials on WEST Using Reciprocating Emissive Probes

An extensive documentation of ICRF-enhanced plasma potentials has been conducted over two experimental campaigns on the WEST tokamak using reciprocating emissive probes magnetically connected to two ICRF antennas. The collected data spans a wide range of antenna electrical settings (coupled power, toroidal phasing, left–right power balance) and plasma parameters (density at the antenna limiter above and below the lower hybrid resonance, plasma current, minority fraction). By scanning the edge safety factor across multiple probe plunges, the magnetic connection between the probe and the antenna varied, enabling the construction of a 2D map of the plasma and floating potentials around an active ICRF antenna. This dataset will be used to validate RF simulation tools equipped with the sheath boundary condition and used to predict RF rectified potentials and ICRF-induced impurity sputtering in future machines. This paper presents the diagnostic and some initial measurements, while the rest will be reported elsewhere.

Diab, Raymond [Massachusetts Inst. of Technology (↗

Spatiotemporal Dynamics of the Relative Abundance of Soil Nutrient‐Degrading Enzyme‐Encoding Genes Across Continental US Ecoregions

Understanding the spatiotemporal patterns in the relative abundance of soil extracellular enzyme‐encoding genes is critical for predicting microbial responses to environmental change and their potential role in nutrient cycling. Yet, integrating novel metagenomic observations with spatiotemporal environmental gradients to infer regional patterns and future trajectories has remained unclear. To address this gap, we applied a machine learning (ML) approach, integrating soil metagenomic data with environmental variables—soil properties, topography, vegetation, and climate—to predict the relative abundance of enzyme‐encoding genes for soil carbon (C), nitrogen (N), and phosphorus (P) across surface soils of the continental United States. We assessed potential responses under future emission scenarios (SSP2‐4.5 and SSP5‐8.5) by comparing a baseline (1985–2014) to a future period (2071–2100). The ML model explained 57%–63% of baseline variation. Precipitation was identified as the most influential factor for the relative abundance of C‐ and N‐degrading enzyme‐encoding genes, while slope length, representing horizontal distance that water can travel downslope, was the primary driver for P‐degrading enzyme‐encoding genes abundance. Projections revealed spatially heterogeneous shifts across continental US ecoregions: the relative abundance of C‐ and N‐degrading enzyme‐encoding genes decreased in wetter ecoregions and increased in drier ecoregions under future climate, while P‐degrading enzyme‐encoding genes abundance decreased significantly in semiarid and Mediterranean ecoregions. This study demonstrates the utility of metagenomic data for mapping soil genetic potential and predicting its regional response to environmental change, to inform ecosystem management strategies.

extracellular enzyme-encoding genes↗

Recurrent features of amplitudes in planar $\mathcal{N}$ = 4 super Yang-Mills theory

The planar three-gluon form factor for the chiral stress tensor operator in planar maximally supersymmetric Yang-Mills theory is an analog of the Higgs-to-three-gluon scattering amplitude in QCD. The amplitude (symbol) bootstrap program has provided a wealth of high-loop perturbative data about this form factor, with results up to eight loops available. The symbol of the form factor at L loops is given by words of length 2L in six letters with associated integer coefficients. In this paper, we analyze this data, describing patterns of zero coefficients and relations between coefficients. We find many sequences of words whose coefficients are given by closed-form expressions which we expect to be valid at any loop order. Moreover, motivated by our previous machine-learning analysis, we identify simple recursion relations that relate the coefficient of a word to the coefficients of particular lower-loop words. These results open an exciting door for understanding scattering amplitudes at all loop orders.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Scalable Hybrid Learning Techniques for Scientific Data Compression

Data compression is becoming critical for storing scientific data because many scientific applications need to store large amounts of data and post process this data for scientific discovery. Unlike image and video compression algorithms that limit errors to primary data (PD), scientists require compression techniques that accurately preserve derived quantities of interest (QoIs). Here, this article presents a physics-informed compression technique implemented as an end-to-end, scalable, GPU-based pipeline for data compression that addresses this requirement. Our hybrid compression technique combines machine learning techniques and standard compression methods. Specifically, we combine an autoencoder, an error-bounded lossy compressor to provide guarantees on raw data error, and a constraint satisfaction post-processing step to preserve the QoIs within a minimal error (generally less than floating point error). The effectiveness of the data compression pipeline is demonstrated by compressing nuclear fusion simulation data generated by a large-scale fusion code, XGC, which produces hundreds of terabytes of data in a single day. Our approach works within the ADIOS framework and results in compression by a factor of more than 150 while requiring only a few percent of the computational resources necessary for generating the data, making the overall approach highly effective for practical scenarios.

ITER↗

Identification of short-range ordering motifs in semiconductors

Chemical short-range ordering is expected to be a key factor for tuning the electronic structure of semiconductors. However, experimental evidence of short-range ordering is still lacking due to the challenge of characterizing atomic-scale ordering motifs. Here, we determined the presence of short-range order in a ternary GeSiSn semiconductor system using advanced energy-filtered four-dimensional scanning transmission electron microscopy and large-scale atomistic models generated by a machine learning neuroevolution potential of first-principles accuracy. This approach revealed preferred ordering of different atomic species with the dominant occurrence of Si–Ge–Sn triplets. Our findings not only confirmed the presence of short-range order but also directly revealed the actual atomic structure, demonstrating the potential for informed atomic order–based band engineering as a third degree of freedom beyond composition and strain tuning.

Vogl, Lilian M. [University of California Berkeley↗

Evaluating Machine Learning-Based MRI Reconstruction Using Digital Image Quality Phantoms

Quantitative and objective evaluation tools are essential for assessing the performance of machine learning (ML)-based magnetic resonance imaging (MRI) reconstruction methods. However, the commonly used fidelity metrics, such as mean squared error (MSE), structural similarity (SSIM), and peak signal-to-noise ratio (PSNR), often fail to capture fundamental and clinically relevant MR image quality aspects. To address this, we propose evaluation of ML-based MRI reconstruction using digital image quality phantoms and automated evaluation methods. Our phantoms are based upon the American College of Radiology (ACR) large physical phantom but created in k-space to simulate their MR images, and they can vary in object size, signal-to-noise ratio, resolution, and image contrast. Our evaluation pipeline incorporates evaluation metrics of geometric accuracy, intensity uniformity, percentage ghosting, sharpness, signal-to-noise ratio, resolution, and low-contrast detectability. We demonstrate the utility of our proposed pipeline by assessing an example ML-based reconstruction model across various training and testing scenarios. The performance results indicate that training data acquired with a lower undersampling factor and coils of larger anatomical coverage yield a better performing model. The comprehensive and standardized pipeline introduced in this study can help to facilitate a better understanding of the performance and guide future development and advancement of ML-based reconstruction algorithms.

47 OTHER INSTRUMENTATION↗

Radiation induced non-linear oscillations in ITER baseline scenario plasmas in DIII-D

Abstract This work shows how the radiation brought about by metals or metal-equivalent radiators such as Kr and Xe produces non-linear dynamics on otherwise stationary β N flattops of DIII-D ITER Baseline Scenario demonstration discharges. The Kr and Xe gases are used to reproduce the radiative loss rates of W in present machines that operate at core temperatures much lower than the expected ITER temperature. Experiments on DIII-D with injection of Kr and Xe, as well as with sources of intrinsic metals reach the range of radiated fraction values expected in the ITER core and experience slow oscillations in temperature and radiated power. In many cases of high radiated fraction, the core temperature decreases enough for the safety factor profile to rise above the 1/1 rational surface, naturally eliminating sawteeth and occasionally producing a persistent helical core. The oscillations can be reproduced by a modified Lotka–Volterra system for temperature and radiated fraction if diffusion and noise are included, which indicates that the interplay between temperature and radiation can be the main cause of the cyclic nature of the system. A new physics based model which includes equations for temperature, density and input power can also reproduce the oscillations observed in the experiments. The present results suggest that the non-linearity of the system can be increased by the inclusion of the inherently non-linear alpha heating term, which is proportional to ∼ n e 2 T i 2 , and obtains oscillations in the model when added to an otherwise more stationary system.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Explaining Health Risk Behaviors in the U.S. with Social Deprivation at Local and Regional Levels

Health risk behaviors are precursors to many chronic health outcomes, and hence, they pose a challenge to public health. Social deprivation undoubtedly creates circumstances that limit access to healthy habits. Moreover, broad regional effects (weather patterns, political ideology, social norms), and local characteristics (cultural notions and barriers, urban places) also influence lifestyle choices and must be accounted for to truly understand the impact of social deprivation on risky behaviors. This research fills the knowledge gap in epidemiological modeling of health risk behaviors by leveraging machine learning to find associations between social deprivation and health risk behaviors, when adjusted by regional and local effects. Four health risk behaviors, namely, binge drinking, smoking, lack of sleep, and lack of physical activity from the CDC PLACES project are considered in a single framework to understand and compare the interplay between local/regional characteristics and seven measures of social deprivation. Our results indicate that local and/or regional factors rise to the top for three out of four risk behaviors (binge drinking, smoking and lack of sleep) out-competing social deprivation measures. Un-entangling the geographical effects reveals that poverty, educational attainment and non-employment are the three deprivation measures most significantly associated with all four health risk factors. The research thus indicates that public health policies to promote healthy lifestyle behaviors must seek to remedy social deprivation, but using socially and culturally sensitive interventions.

Gokhale, Swapna↗

Improvement and generalization of ABCD method with Bayesian inference

To find New Physics or to refine our knowledge of the Standard Model at the LHC is an enterprise that involves many factors, such as the capabilities and the performance of the accelerator and detectors, the use and exploitation of the available information, the design of search strategies and observables, as well as the proposal of new models. We focus on the use of the information and pour our effort in re-thinking the usual data-driven ABCD method to improve it and to generalize it using Bayesian Machine Learning techniques and tools. We propose that a dataset consisting of a signal and many backgrounds is well described through a mixture model. Signal, backgrounds and their relative fractions in the sample can be well extracted by exploiting the prior knowledge and the dependence between the different observables at the event-by-event level with Bayesian tools. We show how, in contrast to the ABCD method, one can take advantage of understanding some properties of the different backgrounds and of having more than two independent observables to measure in each event. In addition, instead of regions defined through hard cuts, the Bayesian framework uses the information of continuous distribution to obtain soft-assignments of the events which are statistically more robust. To compare both methods we use a toy problem inspired by pp\to hh\to b\bar b b \bar b p p → h h → b b ‾ b b ‾ , selecting a reduced and simplified number of processes and analysing the flavor of the four jets and the invariant mass of the jet-pairs, modeled with simplified distributions. Taking advantage of all this information, and starting from a combination of biased and agnostic priors, leads us to a very good posterior once we use the Bayesian framework to exploit the data and the mutual information of the observables at the event-by-event level. We show how, in this simplified model, the Bayesian framework outperforms the ABCD method sensitivity in obtaining the signal fraction in scenarios with 1% and 0.5% true signal fractions in the dataset. We also show that the method is robust against the absence of signal. We discuss potential prospects for taking this Bayesian data-driven paradigm into more realistic scenarios.

Alvarez, Ezequiel↗

Estimators and Fusers for Fiber Delay Estimation Using Environmental Measurements

The properties of deployed network fiber are affected by environmental factors due to their exposure to the elements. Particularly for quantum networks, the resultant delay variations may have significant impacts due to the extreme sensitivity of synchronization, coincidence counting, and other critical operations. In this paper, the delays of 15 km aerial-inground fiber connections are measured, and effects due to temperature, humidity and wind speed are analyzed over multiple periods spanning four seasons of a year. Machine learning methods are first utilized to reveal surprisingly pronounced effects of humidity on the delay, in addition to the expected temperature and its seasonal variations. Estimator and fusion methods are developed to estimate the delay using temperature, humidity and wind speed measurements, by utilizing smooth Gaussian Process Regression (GPR) and nonsmooth Ensemble of Trees (EOT) methods. Measurements from winter and summer periods are temporally fused using twelve different methods, and eight methods provide estimates for the delay throughout the year with median test errors under 1.28%. The results reveal distinct temperature-humidity trends across the seasons, and the ability of estimator and temporal fusion methods to exploit them for estimating the delay. These results constitute a case study of machine learning analytical results, wherein generalization equations explain the performance of various estimator and fuser methods.

Rao, Nageswara [ORNL] (ORCID:0000000234085941)↗

Self-consistent equilibrium and transport simulations for NSTX-U plasmas enhanced via machine learning surrogate models

The Control-Oriented Transport SIMulator (COTSIM) is an advanced equilibrium and transport code designed for simulating tokamak discharges at computational speeds suitable for control applications. COTSIM’s modular framework enables users to select models that balance accuracy with speed according to specific needs, allowing the code to operate from fast to faster-than-real-time performance levels. This work presents recent enhancements to COTSIM’s predictive accuracy for NSTX-U scenarios, achieved by integrating neural-network-based surrogate models and self-consistent equilibrium calculations. To improve source deposition predictions, a surrogate model for NUBEAM has been incorporated. Additionally, a surrogate model for the Multi-Mode Module (MMM) now supports predictions of anomalous thermal, momentum, and particle diffusivities—key factors for modeling the evolution of temperature and rotation. Each surrogate model was specifically trained for the NSTX-U operational regime to enhance COTSIM’s accuracy while maintaining computational efficiency. Moreover, COTSIM now couples fixed-boundary equilibrium solvers with its transport solvers, enabling self-consistent predictions of plasma profiles and equilibrium evolution over the discharge. Simulation results demonstrate strong agreement between COTSIM and TRANSP predictions for NSTX-U discharges. These substantial advancements expand COTSIM’s utility in model-based control applications for NSTX-U. Potential applications include simultaneous optimization of equilibrium and transport scenarios, integration into digital twins, real-time profile estimation (e.g., temperature and rotation) from limited or noisy measurements, and advanced feedback-based scenario control.

Equilibrium and transport modeling↗

Illuminating the Material World: Autonomous Microscopy to Understand Order, Disorder, and Everything In Between

Artificial intelligence (AI) holds immense promise for revolutionizing microscopy, yet its widespread adoption has been hindered by challenges ranging from user inexperience to limited model transferability and difficulties in operationalizing machine learning. This presentation showcases our approach to developing practical autonomy for materials discovery, aiming to accelerate the integration of AI into everyday microscopy workflows. As shown in Fig. 1, I will focus on three key areas: understanding order-disorder transitions, quantifying point defects, and achieving truly device-scale microscopy. First, I will demonstrate the power of multi-modal knowledge graphs for integrating diverse microscopy data. By combining imaging, spectroscopy, and diffraction data, these graphs provide a holistic view of material behavior, capturing the intricate relationships between different modalities [1,2]. I will present a case study on how these models illuminate the structural and chemical changes associated with irradiation in oxide thin films, revealing critical insights for designing materials for extreme environments like spaceflight and nuclear energy. Specifically, I will show how multi-modal analysis clarifies the evolution of order-disorder transitions under irradiation, a key factor influencing material performance in these applications. Next, I will address the challenge of quantifying point defects in 2D materials. We demonstrate the application of computer vision and transfer learning to accurately identify and classify various defect types, such as vacancies and substitutional atoms, and to quantify their concentrations. This information is crucial for understanding and tailoring the properties of 2D materials for applications in electronics, optoelectronics, and catalysis. For example, I will show how our models can characterize the topological distribution of point defects in MXene transition metal carbides, providing valuable insights for optimizing their performance in energy storage and separation science. Finally, I will discuss our progress toward autonomous device-scale microscopy [3,4]. We are fundamentally redesigning electron microscopes around the principles of machine reasoning, enabling automation beyond basic tasks like sample navigation and data acquisition to include sophisticated experimental design. This approach paves the way for truly reproducible and massively scaled analysis campaigns. I will emphasize the importance of autonomous microscopy platforms for high-throughput materials discovery and characterization, facilitating the rapid screening of materials for a broad range of applications and accelerating the development of next-generation technologies.

36 MATERIALS SCIENCE↗