Search NASA⌕ Search

SEARCH · Search NASA

Results for “factorization machine”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 199 records · Page 11

Search for light long-lived particles decaying to displaced jets in proton–proton collisions at $\sqrt{s} = 13.6$ TeV

A search for light long-lived particles (LLPs) decaying to displaced jets is presented, using a data sample of proton–proton collisions at a center-of-mass energy of 13.6 TeV, corresponding to an integrated luminosity of 34.7 fb −1 , collected with the CMS detector at the CERN LHC in 2022. Novel trigger, reconstruction, and machine-learning techniques were developed for and employed in this search. After all selections, the observations are consistent with the background predictions. Limits are presented on the branching fraction of the Higgs boson to LLPs that subsequently decay to quark pairs or tau lepton pairs. An improvement by up to a factor of 10 is achieved over previous limits for models with LLP masses smaller than 60 GeV and proper decay lengths smaller than 1 m. The first constraints are placed on the fraternal twin Higgs (FTH) and folded supersymmetry (FSUSY) models, where the lower bounds on the top quark partner mass reach up to 350 GeV for the FTH model and 250 GeV for the FSUSY model.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Cooperation in Transmission Expansion Planning: Enhancing Grid Reliability and Efficiency Under a Changing Climate

Electricity grids are challenged to maintain reliability during more intense and frequent extreme weather events due to climate change. This challenge is exacerbated by multi-sector electrification and power sector decarbonization through increased reliance on variable renewable energy, which necessitates the expansion of transmission infrastructure. However, transmission expansion planning is often complicated by intertwined planning authorities and jurisdictions, and allocation of large capital investment needs. These factors cause authorities to manage transmission investments individually (i.e., only/mostly intraregional planning), which can lead to suboptimal transmission networks. This study investigates the potential benefits of cooperative transmission expansion planning (i.e., both intraregional and interregional planning that optimizes transmission investments across the entire physical system). Using sectoral and economic optimization, and machine learning models, it analyzes the impact of different levels of cooperation among transmission planning regions within U.S. Western Interconnection in 2019 and 2059 via an iterative investment process. Furthermore, it examines the effects of future climate change on transmission cooperation by simulating historical heat waves from 2019 under conditions of 2059. The results indicate that cooperative transmission planning leads to lower wholesale electricity prices, decreased energy outages, and reduced greenhouse gas emissions. However, the advantages of collaboration diminish during widespread heat waves, despite remaining beneficial especially for regions like California Independent System Operator with substantial solar installations. The study underscores the importance of transmission cooperation in reducing costs and enhancing reliability, emphasizing the need for strategic investments in storage to address challenges posed by future extreme weather events with varying spatial scales.

Capacity Expansion Model↗

Information-entropy-driven generation of material-agnostic datasets for machine-learning interatomic potentials

In contrast to their empirical counterparts, machine-learning interatomic potentials (MLIAPs) promise to deliver near-quantum accuracy over broad regions of configuration space. However, due to their generic functional forms and extreme flexibility, they can catastrophically fail to capture the properties of novel, out-of-sample configurations, making the quality of the training set a determining factor, especially when investigating materials under extreme conditions. We propose a novel automated dataset generation method based on the maximization of the information entropy of the feature distribution, aiming at an extremely broad coverage of the configuration space in a way that is agnostic to the properties of specific target materials. The ability of the dataset to capture unique material properties is demonstrated on a range of unary materials, including elements with the FCC (Al), BCC (W), HCP (Be, Re and Os), graphite (C), and trigonal (Sb, Te) ground states. MLIAPs trained to this dataset are shown to be accurate over a range of application-relevant metrics, as well as extremely robust over very broad swaths of configurations space, even without dataset fine-tuning or hyper-parameter optimization, making the approach extremely attractive to rapidly and autonomously develop general-purpose MLIAPs suitable for simulations in extreme conditions.

36 MATERIALS SCIENCE↗

Deep Learning Advances Arctic River Water Temperature Predictions

The accelerated warming in the Arctic poses serious risks to freshwater ecosystems by altering streamflow and river thermal regimes. However, limited research on Arctic River water temperatures exists due to data scarcity and the absence of robust methodologies, which often focus on large, major river basins. To address this, we leveraged the newly released, extensive AKTEMP data set and advanced machine learning techniques to develop a Long Short-Term Memory (LSTM) model. By incorporating ERA5-Land reanalysis data and integrating physical understanding into data-driven processes, our model advanced river water temperature predictions in ungauged, snow- and permafrost-affected basins in Alaska. Our model outperformed existing approaches in high-latitude regions, achieving a median Nash-Sutcliffe Efficiency of 0.95 and root mean squared error of 1.0°C. The LSTM model learned air temperature, soil temperature, solar radiation, and thermal radiation—factors associated with energy balance—were the most important drivers of river temperature dynamics. Soil moisture and snow water equivalent were highlighted as critical factors representing key processes such as thawing, melting, and groundwater contributions. Glaciers and permafrost were also identified as important covariates, particularly in seasonal river water temperature predictions. Our LSTM model successfully captured the complex relationships between hydrometeorological factors and river water temperatures across varying timescales and hydrological conditions. This scalable and transferable approach can be potentially applied across the Arctic, offering valuable insights for future conservation and management efforts.

54 ENVIRONMENTAL SCIENCES↗

Transforming jet flavour tagging at ATLAS

Jet flavour tagging enables the identification of jets originating from heavy-flavour quarks in proton–proton collisions at the Large Hadron Collider, playing a critical role in its physics programmes. This paper presents GN2, a transformer-based flavour tagging algorithm deployed by the ATLAS Collaboration that represents a different methodology compared to previous approaches. Designed to classify jets based on the flavour of their constituent particles, GN2 processes low-level tracking information in an end-to-end architecture and incorporates physics-informed auxiliary training objectives to enhance both interpretability and performance. Its performance is validated in both simulation and collision data. The measured c-jet (light-jet) rejection in data is improved by a factor of 3.5 (1.8) for a 70% b-jet tagging efficiency, compared to the previous algorithm. GN2 provides substantial benefits for physics analyses involving heavy-flavour jets, such as measurements of Higgs boson pair production and the couplings of bottom and charm quarks to the Higgs boson, and demonstrates the impact of advanced machine learning methods in experimental particle physics.

Characterization and analytical techniques↗

Optimal control of the electron temperature profile in DIII-D using machine learning surrogate models

The viability of the tokamak as a potential fusion reactor depends on the ability to keep the plasma in a stable regime while achieving temperatures, densities, and confinement times that are as high as possible. Tokamak scenario development attempts to find plasma regimes that achieve all of these conditions and are accessible with a given set of hardware constraints. This requires the ability to control plasma properties such as the normalized beta, the internal inductance, safety factor, rotation, etc. One property that has received less attention than some of the others, but is no less critical to achieving high performance, is the electron temperature (T e ) profile. In this work, Linear Quadratic Integral (LQI) control is used to develop a controller for the electron temperature profile in DIII-D. The controller is based on a linearized model derived from the transport equation that describes the evolution of the electron temperature, and includes contributions from the neural network surrogate models NubeamNet and MMMnet. Furthermore, the controller is tested in simulation using COTSIM, and is proven capable of tracking a target T e profile.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Robust Iterative Method for Symmetric Quantum Signal Processing in All Parameter Regimes

Here, this paper addresses the problem of solving nonlinear systems in the context of symmetric quantum signal processing (QSP), a powerful technique for implementing matrix functions on quantum computers. Symmetric QSP focuses on representing target polynomials as products of matrices in SU(2) that possess symmetry properties. We present a novel Newton’s method tailored for efficiently solving the nonlinear system involved in determining the phase factors within the symmetric QSP framework. Our method demonstrates rapid and robust convergence in all parameter regimes, including the challenging scenario with ill-conditioned Jacobian matrices, using standard double precision arithmetic operations. For instance, solving symmetric QSP for a highly oscillatory target function α cos(1000x) (polynomial degree ≈ 1433) takes 6 iterations to converge to machine precision when α = 0.9, and the number of iterations only increases to 18 iterations when α = 1 – 10 -9 with a highly ill-conditioned Jacobian matrix. Leveraging the matrix product state structure of symmetric QSP, the computation of the Jacobian matrix incurs a computational cost comparable to a single function evaluation. Moreover, we introduce a reformulation of symmetric QSP using real-number arithmetics, further enhancing the method’s efficiency. Extensive numerical tests validate the effectiveness and robustness of our approach, which has been implemented in the QSPPACK software package.

97 MATHEMATICS AND COMPUTING↗

Novel ceramic capacitors with ultrahigh energy density and efficiency (Final Technical Report)

Antiferroelectric ceramics are a special class of material that have shown great potential as the dielectric in electrical capacitors due to their high energy- and power-density. During each charge-discharge cycle, the ceramic undergoes transformation to a ferroelectric phase and resumes its antiferroelectric phase. The hysteresis associated with the transitions leads to a mediocre energy efficiency and service lifetime of antiferroelectric capacitors and, hence, their almost absence in commercial products. Under the support of this research project, we first formulated a universal lattice-compatibility theory that included electrostatic polarization energy along with elastic energy and thermal energy to understand the origin of the hysteresis in antiferroelectric oxides. Guided by this compatibility theory, we conducted high-throughput density functional theory (DFT) calculations to assess chemical modifiers and their effect on crystal structures of 400+ PbZrO 3 -based compositions. Down-selected compositions were experimentally validated for their suppressed hysteresis and higher energy efficiency. The verified low-hysteresis compositions were then expanded to an antiferroelectric ceramic library with nearly 500 new compositions (more than 1,500 samples) using high-throughput experiments involving ceramic synthesis and property screening. The large quantity of data generated (theory and experimental) in these tasks were processed by machine-learning techniques and identified trends were fed to the next iteration. In the end, we successfully discovered four compositions with near-zero hysteresis, yielding a world-record energy efficiency of 98.2% at an energy density of 3.0 J/cm 3 . Furthermore, our antiferroelectric ceramic capacitor reaches 79.5 million charge-discharge cycles lifetime, a factor of 80 enhancement over previous antiferroelectric ceramics with large hysteresis. These research accomplishments have not only met the milestones set in the SOPO, but also led to two patent filings, three journal publications (one of them was in Advanced Materials, impact factor 29.4), and nine oral presentations at various venues. Through the course of the project, three postdocs, four Ph.D. students, and one M.S. student were trained. In short, our project established a new methodology in searching next-generation functional ceramics on the fundamental side and discovered several high-efficiency antiferroelectric compositions for capacitors on the applied side. Once fabricated into the multilayer form for commercial applications, these ceramic capacitors can potentially enable the high temperature high power density DC-link capacitors that are critical for the next generation inverters in electric vehicles. The project also significantly contributed to the nation’s workforce development in the STEM fields.

36 MATERIALS SCIENCE↗

NSTXU Diagnostic Disruption Dynamic Loading Represented by Response Spectra

This article presents the results of transient dynamic simulations of loads due to disruption eddy currents on the NSTXU vacuum vessel. Dynamic loading at diagnostic mounting locations is expressed as response spectra derived from the time history results of the dynamic structural simulations of a variety of disruption scenarios. The disruption simulations draw on a history of the project assessments of worst case disruptions for specific components. Major efforts to assess disruption loading have included the vacuum vessel which is the major structural support for the machine, as well as the passive plates (PPs), high harmonic fast wave (HHFW) antenna, and centerstack casing. Each one of these efforts included transient electromagnetic simulations producing time-dependent eddy current Lorentz loads (and in some cases halo loads) which then were applied to time-dependent structural dynamic analyses intended to obtain the proper dynamic amplification factors. In some instances, the EM model and structural model were identical allowing direct transfer of EM forces to the structural model. In other cases, the EM and structural model were not identical and the vector potential (VP) transfer method was used. The results files from these analyses were available (or re-run) to post process in ANSYS Classic time history postprocessor. In conclusion, the ANSYS command is used to create response spectra from time history data at desired points on the vessel.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Evidence of short chains in liquid sulfur

High energy x-ray pair distribution function measurements show the average coordination number of the first shell in liquid sulfur is 1.86 ± 0.04 across the λ-transition, not precisely 2.0 as widely accepted. This indicates that upon melting, liquid sulfur does not comprise solely of S 8 rings but also possesses a significant number of short chains. Intensities of the pre-peak and first diffraction peak of the x-ray structure factor and third peak height of the pair distribution function all show deviations at the λ-transition temperature T λ , associated with the break-up of S 8 rings and the start of oligomer polymerization. A significant number of non-bonded or loosely bonded “interstitial atoms,” with an average coordination number of 0.20 ± 0.005, are also observed in the so-called “forbidden zone” between the first and second shells upon melting. The number of interstitial atoms is found to decrease to a minimum at the λ-transition, but the majority persist into the high temperature polymerized liquid. Furthermore, the existence of short chains and nearby interstitial atoms represent the two main factors required to initiate the S 8 -ring to chain transition, as proposed by recent molecular dynamics simulations.

Chemical bonding↗

Integrating Crack Detection and Pipe Shape Optimization for Enhanced Sewage System Durability

Crack detection in underground reinforced concrete pipes has been essential in determining the state of stormwater infrastructure. Detection models have been implemented for detecting cracks and other defects in pipes using CCTV footage for stormwater drainage systems. In addition, Finite element models have been used to determine optimum shapes and pipe thickness for different boundary conditions such as header pipes in power plants. The concept of shape optimization emerges as a crucial factor in power plant design and operation, with the potential to maximize performance while minimizing the use of materials. Shape optimization not only enhances efficiency but also contributes to reducing the environmental footprint. This paper discusses the integration of both topics by using the cracks detected in underground pipes as boundary conditions for shape optimization of the pipes. A machine learning model has been developed which uses limited data for training and outlines the location of detected cracks. A shape optimization methodology is proposed in which ANSYS modules are used to analyze fluid flow and then optimize the shape of the pipe. The crack detection model developed has been applied to a crack detected in lab setting and machine learning model used has an accuracy of 98% using a random forest algorithm.

20 FOSSIL-FUELED POWER PLANTS↗

Separable physics-informed DeepONet: Breaking the curse of dimensionality in physics-informed machine learning

The deep operator network (DeepONet) has shown remarkable potential in solving partial differential equations (PDEs) by mapping between infinite-dimensional function spaces using labeled datasets. However, in scenarios lacking labeled data, the physics-informed DeepONet (PI-DeepONet) approach, which utilizes the residual loss of the governing PDE to optimize the network parameters, faces significant computational challenges, particularly due to the curse of dimensionality. This limitation has hindered its application to high-dimensional problems, making even standard 3D spatial with 1D temporal problems computationally prohibitive. Additionally, the computational requirement increases exponentially with the discretization density of the domain. Here, to address these challenges and enhance scalability for high-dimensional PDEs, we introduce the Separable physics-informed DeepONet (Sep-PI-DeepONet). This framework employs a factorization technique, utilizing sub-networks for individual one-dimensional coordinates, thereby reducing the number of forward passes and the size of the Jacobian matrix required for gradient computations. By incorporating forward-mode automatic differentiation (AD), we further optimize computational efficiency, achieving linear scaling of computational cost with discretization density and dimensionality, making our approach highly suitable for high-dimensional PDEs. We demonstrate the effectiveness of Sep-PI-DeepONet through three benchmark PDE models: the viscous Burgers’ equation, Biot’s consolidation theory, and a parameterized heat equation. Our framework maintains accuracy comparable to the conventional PI-DeepONet while reducing training time by two orders of magnitude. Notably, for the heat equation solved as a 4D problem, the conventional PI-DeepONet was computationally infeasible (estimated 289.35 h), while the Sep-PI-DeepONet completed training in just 2.5 h. These results underscore the potential of Sep-PI-DeepONet in efficiently solving complex, high-dimensional PDEs, marking a significant advancement in physics-informed machine learning.

Neural operator↗

Machine learning-accelerated path integral molecular dynamics simulations of reactive organic electrolytes

Hydrogen bonded electrolytes that exhibit accelerated proton transport via sequential reactive hops have drawn interest for their promise in clean energy applications. Molecular dynamics simulations of these electrolytes offer the opportunity to uncover microscopic mechanistic details that could be used to design and tune the properties of candidate electrolyte technologies. However, accurately modeling the proton transfer reactions and transport properties that give rise to high charge conductivites in these electrolytes proves computationally challenging because of the need to perform lengthy condensed phase simulations, treating both the electronic and nuclear degrees of freedom quantum mechanically. In this paper, we demonstrate that such a modeling task can be efficiently achieved with the use of density functional theory (DFT)-trained machine learning potentials (MLP) to accelerate path integral molecular dynamics (PIMD) simulations. We highlight the practical utility of this approach by using it to benchmark how closely PIMD simulations employing different DFT exchange–correlation functionals reproduce the composition-dependent densities, diffusion coefficients, and electrical conductivities of mixtures consisting of imidazole and levulinic acid. Even with the speedup afforded by our MLPs, PIMD simulations remain quite expensive. Furthermore, in order to render PIMD more computationally tractable, we introduce and benchmark the accuracy of a ring polymer contraction approach that leverages a computationally efficient short-range MLP to accelerate our PIMD simulations by an additional factor of four.

Chemical bonding↗

wa-hls4ml: A Benchmark and Surrogate Models for hls4ml Resource and Latency Estimation

As machine learning (ML) is increasingly implemented in hardware to address real-time challenges in scientific applications, the development of advanced toolchains has significantly reduced the time required to iterate on various designs. These advancements have solved major obstacles, but also exposed new challenges. For example, processes that were not previously considered bottlenecks, such as hardware synthesis, are becoming limiting factors in the rapid iteration of designs. To mitigate these emerging constraints, multiple efforts have been undertaken to develop an ML-based surrogate model that estimates resource usage of ML accelerator architectures. We introduce wa-hls4ml, a benchmark for ML accelerator resource and latency estimation, and its corresponding initial dataset of over 680,000 fully connected and convolutional neural networks, all synthesized using hls4ml and targeting Xilinx FPGAs. The benchmark evaluates the performance of resource and latency predictors against several common ML model architectures, primarily originating from scientific domains, as exemplar models, and the average performance across a subset of the dataset. Additionally, we introduce GNN- and transformer-based surrogate models that predict latency and resources for ML accelerators. We present the architecture and performance of the models and find that the models generally predict latency and resources for the 75% percentile within several percent of the synthesized resources on the synthetic test dataset.

Hawks, Benjamin [Fermilab] (ORCID:0000000157000288↗

wa-hls4ml: A Benchmark and Surrogate Models for hls4ml Resource and Latency Estimation

As machine learning (ML) is increasingly implemented in hardware to address real-time challenges in scientific applications, the development of advanced toolchains has significantly reduced the time required to iterate on various designs. These advancements have solved major obstacles, but also exposed new challenges. For example, processes that were not previously considered bottlenecks, such as hardware synthesis, are becoming limiting factors in the rapid iteration of designs. To mitigate these emerging constraints, multipleefforts have been undertaken to develop an ML-based surrogate model that estimates resource usage of synthesized ML accelerator architectures. We introduce wa-hls4ml, a benchmark for ML accelerator resource and latency estimation, and its corresponding initial dataset of over 680 000 fully connected and convolutional neural networks, all synthesized using hls4ml and targeting Xilinx FPGAs. The benchmark evaluates the performance of resource and latency predictors against several common ML model architectures, primarily originating from scientific domains, as exemplar models, and the average performance across a subset of the dataset. Additionally, we introduce GNN- and transformer-based surrogate models that predict latency and resources for ML accelerators. We present the architecture and performance of the models and find that the models generally predict latency and resources for the 75% percentile within several percent of the synthesized resources on the synthetic test dataset.

Hawks, Benjamin G. [Fermilab]↗

Single Photon Searches at ICARUS with SPINE

The MiniBooNE Low-Energy Excess (LEE) of electron-like events from the Booster Neutrino Beam (BNB) has puzzled neutrino physicists for decades. One possible explanation has been an unpredicted excess of neutral current (NC) $\Delta$ resonance interactions with a subsequent radiative decay. An increase in the rate of NC $\Delta\rightarrow N\gamma$ events by a factor of 3.18 could explain the LEE seen by MiniBooNE. ICARUS also sees neutrinos from the BNB and can check the rate of these single photon events. The SPINE particle physics reconstruction suite leverages deep neural networks (DNNs) to optimize particle reconstruction and identification in Liquid Argon Time Projection Chambers (LArTPCs). I present preliminary findings on the effectiveness of these machine learnign (ML) techniques for single photon event reconstruction at ICARUS.

Hausner, Harry [Fermilab] (ORCID:0000000188932280)↗

Machine learning based prediction of airflow maldistribution in air-to-refrigerant heat exchangers

Flow maldistribution is a common challenge in heat exchanger (HX) design and particularly important for air-to-refrigerant geometries where capacity losses can approach 65%. This has a major impact on central air conditioning systems, as compact duct design motivates the use of A-type HXs which are known to be affected by airflow maldistribution. Because velocity profiles are difficult to predict, components are often oversized leading to increased material cost, system footprint, and refrigerant charge. Several studies detail airflow maldistribution for individual HXs and packages, but findings cannot always be extrapolated to new designs. In this work, a machine learning (ML) based flow profile prediction framework is developed and applied to two common package configurations: (i) A-type and (ii) U-type HXs, across a broad range of HX geometries and flow rates. Porous media CFD simulations are validated against independent data for both package types as well as comprehensive in house measurements for a finless geometry with shape optimized non-round tubes, which validates the framework for new heat transfer surfaces. The ML models are trained on the porous media CFD simulations, predicting volumetric flow rate (VFR) within 1.1% and 1.9% with maximum relative L 2 norm errors of 0.48 and 0.65, respectively, while also delivering 10 5 speed up factor compared to full porous media CFD. HX level simulations show an up to 9% reduction in heat transfer from flow maldistribution, with greater losses occurring at smaller half apex angles. This framework enables rapid and highly accurate prediction of airflow maldistribution induced capacity degradation.

42 ENGINEERING↗

Kernel methods for evolution of generalized parton distributions

Generalized parton distributions (GPDs) characterize the 3-dimensional structure of hadrons, combining information about their internal quark and gluon longitudinal momentum distributions and transverse position within the hadron. The dependence of GPDs on the factorization scale Q 2 allows one to connect hard exclusive processes involving GPDs at disparate energy and momentum scales, which is needed in global analyses of experimental data. Here, in this work, we explore how finite element methods can be used to construct fast and differentiable Q 2 evolution codes for GPDs in momentum space, which can be used in a machine learning framework. We show numerical benchmarks of the methods' accuracy, including a comparison to an existing evolution code from PARTONS/APFEL++, and provide a repository where the code can be accessed.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗