Search NASA⌕ Search

SEARCH · Search NASA

Results for “Factorization machine”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 307 records · Page 17

Low-index mesoscopic surface reconstructions of Au surfaces using Bayesian force fields

Metal surfaces have long been known to reconstruct, significantly influencing their structural and catalytic properties. Many key mechanistic aspects of these subtle transformations remain poorly understood due to limitations of previous simulation approaches. Using active learning of Bayesian machine-learned force fields trained from ab initio calculations, we enable large-scale molecular dynamics simulations to describe the thermodynamics and time evolution of the low-index mesoscopic surface reconstructions of Au (e.g., the Au(111)-‘Herringbone,’ Au(110)-(1 × 2)-‘Missing-Row,’ and Au(100)-‘Quasi-Hexagonal’ reconstructions). This capability yields direct atomistic understanding of the dynamic emergence of these surface states from their initial facets, providing previously inaccessible information such as nucleation kinetics and a complete mechanistic interpretation of reconstruction under the effects of strain and local deviations from the original stoichiometry. We successfully reproduce previous experimental observations of reconstructions on pristine surfaces and provide quantitative predictions of the emergence of spinodal decomposition and localized reconstruction in response to strain at non-ideal stoichiometries. A unified mechanistic explanation is presented of the kinetic and thermodynamic factors driving surface reconstruction. Furthermore, we study surface reconstructions on Au nanoparticles, where characteristic (111) and (100) reconstructions spontaneously appear on a variety of high-symmetry particle morphologies.

36 MATERIALS SCIENCE↗

Subsurface hydrological controls on the short-term effects of hurricanes on nitrate–nitrogen runoff loading: a case study of Hurricane Ida using the Energy Exascale Earth System Model (E3SM) Land Model (v2.1)

When the nutrient level in the soil surpasses vegetation demand, nutrient losses due to surface runoff and subsurface leaching are the major reasons for the deterioration of water quality. The lower Mississippi River basin (LMRB) is one of the sub-basins that deliver the highest nitrogen loads to the Gulf of Mexico. Potential changes in episodic events induced by hurricanes may exacerbate water quality issue in the future. However, uncertainties in modeling the hydrologic response to hurricanes may limit the modeling of nutrient losses during such events. Using a machine learning approach, we calibrated the land component of the Energy Exascale Earth System Model (E3SM), or ELM, version 2.1, based on the water table depth (WTD) of a calibrated 3D subsurface hydrology model. While the overall performance of the calibrated ELM is satisfactory, some discrepancies in WTD remain in slope areas with low precipitation due to the missing lateral flow process in ELM. Simulations including biogeochemistry performed using ELM with and without model calibration showed important influences of soil hydrology, precipitation intensity, and runoff parameterization on the magnitude of nitrogen runoff loss and the leaching pathway. Despite such sensitivities, both ELM simulations produced reduced WTD and increased runoff and accelerated nitrate–nitrogen runoff loading during Hurricane Ida in August 2021, consistent with the observations. With observations suggesting more pronounced effects of Hurricane Ida on nitrogen runoff than the simulations, we identified factors for model improvement to provide a useful tool for studying hurricane-induced nutrient losses in the LMRB region.

54 ENVIRONMENTAL SCIENCES↗

Linking large-scale weather patterns to observed and modeled turbine hub-height winds offshore of the US West Coast

The US West Coast holds great potential for wind power generation, although its potential varies due to the complex coastal climate. Characterizing and modeling turbine hub-height winds under different weather conditions are vital for wind resource assessment and management. This study uses a two-stage machine learning algorithm to identify five large-scale meteorological patterns (LSMPs): post-trough, post-ridge, pre-ridge, pre-trough, and California high. The LSMPs are linked to offshore wind patterns, specifically at lidar buoy locations within lease areas for future wind farm development off Humboldt and Morro Bay. While each LSMP is associated with characteristic large-scale atmospheric conditions and corresponding differences in wind direction, diurnal variation, and jet features at the two lidar sites, substantial variability in wind speeds can still occur within each LSMP. Wind speeds at Humboldt increase during the post-trough, pre-ridge, and California-high LSMPs and decrease during the remaining LSMPs. Morro Bay has smaller responses in mean speeds, showing increased wind speed during the post-trough and California-high LSMPs. Besides the LSMPs, local factors, including the land–sea thermal contrast and topography, also modify mean winds and diurnal variation. The High-Resolution Rapid Refresh model analysis does a good job of capturing the mean and variation at Humboldt but produces large biases at Morro Bay, particularly during the pre-ridge and California-high LSMPs. The findings are anticipated to guide the selection of cases for studying the influence of specific large-scale and local factors on California offshore winds and to contribute to refining numerical weather prediction models, thereby enhancing the efficiency and reliability of offshore wind energy production.

17 WIND ENERGY↗

Investigating Characteristic Droplet Size Distributions in Large Eddy Simulations of Stratocumulus Clouds

Cloud processes relevant to radiative and precipitation properties depend on the shape of the cloud droplet size distribution. Recent holographic observations revealed that cloud droplet populations do not have the same size distribution shapes throughout but form regions of characteristic distributions with similar microphysical properties. We investigate the existence and properties of these characteristic distributions within Large‐Eddy Simulations of stratocumulus clouds using Lagrangian and bin microphysics schemes. Distribution types are identified, revealing localized characteristic distributions that vary on the scale of the largest convective cell for simulations with bin microphysics. The results from the Lagrangian microphysics scheme hint at similar behavior. Compared to observations, the simulated clouds are much more uniform. Analysis of the LES results suggests a connection to the local entrainment rate, so the poorly resolved entrainment interface in LES may be a cause of the uniformity. The uniformity of the large‐scale forcing could also be a factor.

cloud droplet size distributions↗

DEPRECATED AI-Batt-OS (Autonomous Identification of Battery Life Models - Open Source) [SWR 21-17]

DEPRECATED. This repository was archived by the owner on Jun 30, 2026. It is now read-only. Open source implementation of some of the methods utilized by AI-Batt, a battery lifetime modeling and analysis toolkit provided by the National Laboratory of the Rockies (NLR). This software demonstrates the use of bi-level optimization and symbolic regression techniques to semi-autonomously identify algebraic models predicting the capacity fade of lithium-ion batteries during calendar aging. Modeling the degradation of batteries is a complex task, due to the difficulty in separating the time-dependent and time-independent factors impacting cell level degradation, across multiple data series with different numbers of measurements and/or data quality. Bi-level optimization enables model parameters to be optimized to either the entire data set or to individual data series, allowing statistical disambiguation of global behaviors (data series independent) and local behaviors (data series dependent). Symbolic regression is used to automatically search for optimal low-dimesional models predicting the variation of locally optimized parameters versus time-independent experimental variables from millions of possible models, resulting in a more accurate and repeatable model identification process than is possible by a manual search. The provided tools also implement cross-validation and bootstrap resampling schemes, empowering statistical model comparison/selection and quantification of model uncertainties. An example script replicates the results from the manuscript "Challenging Practices of Algebraic Battery Life Models through Statistical Validation and Model Identification via Machine-Learning", submitted to ECS. All code is written in MATLAB. Requires the Statistics and Machine Learning Toolbox. Contact Dr. Paul Gasper at Paul.Gasper@nlr.gov for any questions.

Gasper, Paul [National Renewable Energy Lab. (NREL↗

Using convolutional neural networks to accelerate three-dimensional coherent synchrotron radiation computations

Calculating the effects of coherent synchrotron radiation (CSR) is one of the most computationally expensive tasks in accelerator physics. Here, we use convolutional neural networks (CNNs), along with a latent conditional diffusion (LCD) model, trained on physics-based simulations to speed up calculations. Specifically, we produce the 3D CSR wakefields generated by electron bunches in circular orbit in the steady-state condition. Two datasets are used for training and testing the models: wakefields generated by three-dimensional Gaussian electron distributions and wakefields from a sum of up to 25 three-dimensional Gaussian distributions. The CNNs are able to accurately produce the 3D wakefields ∼250–1000 times faster than the numerical calculations, while the LCD achieves a gain of a factor of ∼34. We also test the extrapolation and out-of-distribution generalization ability of the models. They generalize well on distributions with larger spreads than what they were trained on but struggle with smaller spreads.

43 PARTICLE ACCELERATORS↗

Binding profiles for 961 Drosophila and C. elegans transcription factors reveal tissue-specific regulatory relationships

A catalog of transcription factor (TF) binding sites in the genome is critical for deciphering regulatory relationships. Here, we present the culmination of the efforts of the modENCODE (model organism Encyclopedia of DNA Elements) and modERN (model organism Encyclopedia of Regulatory Networks) consortia to systematically assay TF binding events in vivo in two major model organisms,Drosophila melanogaster(fly) andCaenorhabditis elegans(worm). These data sets comprise 605 TFs identifying 3.6 M sites in the fly and 356 TFs identifying 0.9 M sites in the worm, and represent the majority of the regulatory space in each genome. We demonstrate that TFs associate with chromatin in clusters termed “metapeaks,” that larger metapeaks have characteristics of high-occupancy target (HOT) regions, and that the importance of consensus sequence motifs bound by TFs depends on metapeak size and complexity. Combining ChIP-seq data with single-cell RNA-seq data in a machine-learning model identifies TFs with a prominent role in promoting target gene expression in specific cell types, even differentiating between parent–daughter cells during embryogenesis. These data are a rich resource for the community that should fuel and guide future investigations into TF function. To facilitate data accessibility and utility, all strains expressing green fluorescent protein (GFP)-tagged TFs are available at the stock centers for each organism. The chromatin immunoprecipitation sequencing data are available through the ENCODE Data Coordinating Center, GEO, and through a direct interface that provides rapid access to processed data sets and summary analyses, as well as widgets to probe the cell-type-specific TF–target relationships.

Biochemistry & Molecular Biology↗

What regulates decomposition in agroecosystems? Insights from reading the tea leaves

Litter decomposition is a critical Earth process, recycling nutrients and setting a portion of plant tissue on a path toward soil organic matter. Despite this importance, we still lack a good understanding of local factors that regulate decomposition, especially in agroecosystems where management plays an outsized role. Using a narrow range of climate and soils, we buried 1,308 pre-manufactured “litter bags” of differing residue quality (i.e., green and rooibos tea leaves) in 109 plots across several management practices to (1) explore the local controls on decomposition in agroecosystems and (2) test the robustness of the Tea Bag Index (TBI). We found that management practices intended to increase soil ecosystem services, that is, soil health, altered the decomposition of both teas. For example, adding nitrogen fertilizer and implementing perennial cropping decreased the extent of green tea decomposition (carbon-to-nitrogen ratio, or C:N = 12.8). No-tillage increased, but perennial cropping decreased, the rate of rooibos tea decomposition (C:N = 50.1). Cropped prairie accelerated green tea decomposition and increased the extent of red tea decomposition. A random forest regression model showed that soil temperature was the strongest predictor of green tea decomposition, but a soil health score also played a significant role in predicting the mass remaining. Soil texture and nutrient availability best predicted rooibos tea decomposition. Finer textured soils seemed to decelerate rooibos decomposition but increased the extent of decomposition. Furthermore, we demonstrated that the TBI metrics correlated somewhat well with empirically derived decomposition constants and were similarly sensitive to the effects of management. Still, the green tea stabilization factor had a substantial prediction bias. Our study increased our basic understanding of what regulates decomposition in agroecosystems. It also showed that the TBI can be a scientifically rigorous citizen science approach to monitoring changes in soil health.

60 APPLIED LIFE SCIENCES↗

Empirical scaling of the L–H threshold power for metal wall tokamaks using a multi-device database

The empirical scaling for the H-mode power threshold in tokamaks has been revisited using a database with threshold data from machines with a metallic first wall as part of International Tokamak Physics Activity (ITPA) task TC-26. The database contains discharges from ASDEX Upgrade (AUG) (W), JET (Be/W) and Alcator C-Mod (Mo). This was motivated by reports that in like-for-like discharges the power threshold was reduced by approximately 30% after the change from carbon based to metallic first wall materials on AUG (Ryter et al 2013 Nucl. Fusion 53 113003) and JET (Maggi et al 2014 Nucl. Fusion 54 023007). The database contains L–H transition data for all hydrogen isotopes and mixtures, including T and DT from the recent JET campaigns. Compared to the ITPA 2008 scaling (Martin et al 2008 J. Phys.: Conf. Ser. 123 012033), the metal wall scaling has a smaller magnetic field exponent but a larger density exponent. We present an additional parameter to capture the strong dependence of the L–H power threshold (approx. factor 2) on the magnetic configuration in the divertor on JET. The scaling recovers the approximate inverse isotope mass scaling of the threshold power. Alternative scalings involving the plasma current and poloidal magnetic field are explored. Despite the reduction in threshold observed earlier, the scalings based on the metal wall database do not necessarily extrapolate to a lower threshold for ITER compared to the ITPA 2008 scaling, especially at high density. The divertor configuration effect induces the largest uncertainty in the extrapolation.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Understanding drivers of oil and gas well integrity issues in the greater wattenberg area of Colorado

Well integrity is critically important to maintain to minimize the environmental impacts of oil and gas development and other subsurface energy operations. The Wattenberg Field of Colorado—a top producing field with >40,000 wells—has one of the most robust publicly reported well integrity programs in the country. Here, in this study, we analyzed annular pressure and annular-fluid geochemical test results collected from Wattenberg wells through the end of 2019 to characterize the frequency and spatial variability of integrity issues in the field and understand their drivers. Estimated frequencies of integrity issues among tested wells were 8.2-17.1% between 1955 and 2019 and 6.1-11.4% in 2019 alone. The frequency of integrity issues was nearly four times greater in wells located above the Longmont Wrench Fault Zone. Potential drivers of integrity issues were identified using ensemble decision tree models trained with a broad set of relevant information. Models show that well integrity issues are spatially clustered on regional and sub-regional scales and suggest the relatively high frequency of integrity issues observed is likely attributed to geologic factors. These findings are valuable for regulatory agencies and operators seeking to inform well integrity monitoring, plugging, and emissions reduction efforts and design future subsurface energy projects.

03 NATURAL GAS↗

An Integrated Framework for Risk Assessment of Safety-related Digital Instrumentation and Control Systems in Nuclear Power Plants: Methodology Advancement and Application

This report documents activities performed by Idaho National Laboratory (INL) during fiscal year (FY) 2024 for the U.S. Department of Energy (DOE) Light Water Reactor Sustainability (LWRS) Program, Risk Informed Systems Analysis (RISA) Pathway, Digital Instrumentation and Control (DI&C) Risk Assessment project. The goal of the RISA Pathway is to optimize safety margins and minimize uncertainties to achieve economic efficiencies while maintaining high levels of safety. This is accomplished by providing scientific basis to better represent safety margins and factors that contribute to cost and safety, and by developing new technologies that reduce operating costs. The research efforts for FY 2024 encompass methodology refinement and exploration. The efforts include: (1) The implementation of a natural language processing tool to expedite key aspects of the reliability analysis methods developed by INL; (2) advances to support intersystem CCF analysis by providing guidance for and identification of coupling mechanisms that may contribute to CCF; (3) the investigation of how generative artificial intelligence tools can aid in hazard analysis and diversity and defense in depth (i.e., D3) assessments; (4) Industry collaboration, allowing the demonstration of and INL's risk assessment tools to support risk assessment of DI&C systems at early and late stages of development; (4) a roadmap for the development of a software for each of INL's risk assessment tools; (5) The development of a theory and methodology manual for a risk quantification methodology; (6) the development of a reliability analysis for machine learning (ML)-integrated control systems.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Data-driven Community-centered Resilient Assessment and Planning Toolkit for Nexus of Energy and Water (DCRAPT-NEW)

Urban areas, including Detroit and Pittsburgh, have suffered significant dual outages of the electrical and water infrastructure in the past decade due, in part, to the increasing number of extreme weather events. With increasing temperatures and rainfall intensity, these regions need to prepare for increasing extreme events through community-based energy and water resilience analysis, planning, and enhancement. This project developed a suite of open-source, open-access, community-centered, data-driven assessment and distributed energy resource (DER) and planning tools for energy and water resilience enhancement in urban areas. Through establishing a multi-level community awareness and engagement mechanism and a comprehensive collection of power outage and flooding data, an innovative group of community energy and water resilience assessment and planning tools have been developed for a wide range of users with differing and variable sets of data available to them. The developed tools include (1) DOE EAGLE-I data-driven, deep-learning assisted resilience assessment and DER planning tools at the county level with socioeconomic factors incorporated; (2) Utility annual power outage data-driven tools for long term resilience assessment and DER planning and 15-min power outage data-driven tools for short term resilience assessment and planning; (3) Detailed engineering tools for energy and water systems resilience assessment and planning when the system topology and component fragility curves are available; (4) Alternative Resiliency Metric Calculation that extracts and separates outage and restoration processes; and (5) Co-optimization tools that evaluate the resilience of the power and sewage system and allow users to conduct joint planning with energy and wastewater systems. The developed tools provide planners, decision-makers, and stakeholders with powerful capabilities to systematically evaluate system/community resilience and optimal and actionable guidance for enhancing resilience while prioritizing DER investments. The tools have been used and validated in Detroit and Pittsburgh and can be used in other areas of the nation. In addition, this project will (1) advance the knowledge and applications of machine-learning methods in analyzing and fusing different layers of information and generating meaningful data points such as generating rare weather events; (2) significantly improve the energy and water resilience of the identified communities in Detroit and Pittsburgh and prepare for more frequent and severe weather conditions; (3) help communities assess extreme weather event impacts and address short-term and long-term resilience-related issues The developed tools have been made public via GitHub and demonstrated to community stakeholders and utility companies via the two annual workshops and numerous community engagement meetings. The project outcomes are also disseminated through publications in various journals and conference proceedings, and presentations at top conferences.

13 HYDRO ENERGY↗

Study of the Protection Improvements for a Weak Grid Area With High Inverter-Based Resources (IBRs)

This project designs enhanced protection scheme for the real-world weak grid area with a high penetration of IBRs. As the existing protection schemes are originally designed for traditional synchronous machines, we first evaluate if the protection scheme will continue to operate reliably in systems with high levels of IBRs. Hardware relays are tested using a controller-hardware-in-the-loop setup. PSCAD electromagnetic transient simulation with IBR original equipment manufacturer black-box models is used to perform fault studies and generate COMTRADE data, which are replayed by a real-time digital simulator (RTDS) to feed input to the hardware relays. Three scenarios are analyzed: normal operation, an N-1 contingency, and an IBR-only scenario. The evaluation results reveal the following: 1) the protection scheme remains reliable under normal conditions and N-1 contingencies and 2) in IBR-only scenarios, differential protection (87L) continues to operate reliably, whereas local protection elements, such as distance and directional elements, fail because of the lack of regulated negative sequence current contributed by IBRs. Enhanced protection is designed to address the challenge of lack of negative sequence current from IBRs, including increased restraining factors a2 and k2 to block 32Q or using V instead QV ORDER for ground faults, enhanced mho distance element with voltage and phase angle supervision for L-L faults. The efficacy of enhanced protection logic is validated and proven to work reliably. Additionally, IEEE Std. 2800-2022 negative sequence current compliant GFL and GFM IBRs from another vendor are tested and proven to work reliably without need for enhanced logic. Therefore, this work provides valuable decision-making for utilities facing protection system challenges due to IBRs, either designing enhanced protection scheme or requesting their IBRs being IEEE Std. 2800-2022 compliant to produce regulated negative sequence current for protection relay to make correct decision.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Quantifying Streambed Grain Size, Uncertainty, and Hydrobiogeochemical Parameters Using Machine Learning Model YOLO

Abstract Streambed grain sizes control river hydro‐biogeochemical (HBGC) processes and functions. However, measuring their quantities, distributions, and uncertainties is challenging due to the diversity and heterogeneity of natural streams. This work presents a photo‐driven, artificial intelligence (AI)‐enabled, and theory‐based workflow for extracting the quantities, distributions, and uncertainties of streambed grain sizes from photos. Specifically, we first trained You Only Look Once, an object detection AI, using 11,977 grain labels from 36 photos collected from nine different stream environments. We demonstrated its accuracy with a coefficient of determination of 0.98, a Nash–Sutcliffe efficiency of 0.98, and a mean absolute relative error of 6.65% in predicting the median grain size of 20 ground‐truth photos representing nine typical stream environments. The AI is then used to extract the grain size distributions and determine their characteristic grain sizes, including the 10th, 50th, 60th, and 84th percentiles, for 1,999 photos taken at 66 sites within a watershed in the Northwest US. The results indicate that the 10th, median, 60th, and 84th percentiles of the grain sizes follow log‐normal distributions, with most likely values of 2.49, 6.62, 7.68, and 10.78 cm, respectively. The average uncertainties associated with these values are 9.70%, 7.33%, 9.27%, and 11.11%, respectively. These data allow for the computation of the quantities, distributions, and uncertainties of streambed HBGC parameters, including Manning's coefficient, Darcy‐Weisbach friction factor, top layer interstitial velocity magnitude, and nitrate uptake velocity. Additionally, major sources of uncertainty in grain sizes and their impact on HBGC parameters are examined.

58 GEOSCIENCES↗

Evaluating Cloud Properties at Scott Base: Comparing Ceilometer Observations With ERA5, JRA55, and MERRA2 Reanalyses Using an Instrument Simulator

This study compares CL51 ceilometer observations made at Scott Base, Antarctica, with statistics from the ERA5, JRA55, and MERRA2 reanalyses. To enhance the comparison we use a lidar instrument simulator to derive cloud statistics from the reanalyses which account for instrumental factors. The cloud occurrence in the three reanalyses is slightly overestimated above 3 km, but displays a larger underestimation below 3 km relative to observations. Unlike previous studies, we see no relationship between relative humidity and cloud occurrence biases, suggesting that the cloud biases do not result from the representation of moisture. We also show that the seasonal variation of cloud occurrence and cloud fraction, defined as the vertically integrated cloud occurrence, are small in both the observations and the reanalyses. We also examine the quality of the cloud representation for a set of weather states derived from ERA5 surface winds. The variability associated with grouping cloud occurrence based on weather state is much larger than the seasonal variation, highlighting weather state is a strong control of cloud occurrence. All the reanalyses continue to display underestimates below 3 km and overestimates above 3 km for each weather state. But the variability in ERA5 statistics matches the changes in the observations better than the other reanalyses. We also use a machine learning scheme to estimate the quantity of supercooled liquid water cloud from the ceilometer observations. Ceilometer low-level supercooled liquid water cloud occurrences are considerably larger than values derived from the reanalyses, further highlighting the poor representation of low-level clouds in the reanalyses.

54 ENVIRONMENTAL SCIENCES↗

Application of artificial intelligence methods in the international roughness index prediction of rigid and composite pavements: a systematic review

The International Roughness Index (IRI) is a widely adopted metric for quantifying pavement roughness, directly influencing vehicle safety, ride comfort, and overall roadway performance. In recent years, the use of Machine Learning (ML) models for IRI prediction has gained momentum, with the goal of improving the allocation of maintenance and rehabilitation resources by enabling accurate assessments of pavement conditions. Most prior reviews, however, have concentrated on flexible pavements, leaving a notable gap regarding rigid and composite pavements. To address this gap, the present study conducts a systematic review of Artificial Intelligence (AI) methods applied to IRI prediction for rigid and composite pavements. Literature published between 2004 and 2025 is synthesized to highlight prevailing trends, methodological contributions, and directions for future research. Particular attention is given to the types of models employed, the datasets used for training and validation, and the role of input variables and data-processing strategies. Across the included studies, ensemble learning methods (especially gradient boosting variants such as XGBoost), artificial neural networks, and hybrid architectures frequently achieved high predictive skill, with several models reporting test-set coefficients of determination approaching 0.9–0.96, indicating strong potential for capturing the influence of traffic, pavement structure, and climatic factors. Since these results are obtained from heterogeneous datasets and evaluation protocols, they are interpreted qualitatively rather than as strict cross-study rankings. Analysis of input variables revealed that pavement age and initial IRI were included in 91% (21 of 23) and 78% (18 of 23) of studies, respectively. Climatic variables such as the freezing index appeared in 57% (13 of 23), while traffic-related factors were considered in 65% (15 of 23). The findings underscore the importance of standardized, high-quality datasets, such as those from the Long-Term Pavement Performance (LTPP) program, along with data consistency, model interpretability, computational efficiency, and replicability in enhancing IRI prediction. Future research should focus on incorporating input variable selection techniques to identify the most influential predictors, thereby improving accuracy and robustness. Integrating these approaches with advanced non-linear data-driven models, coupled with robust hyperparameter optimization, holds considerable promise for strengthening the reliability of IRI prediction and supporting resilient pavement management strategies.

42 ENGINEERING↗

Multi-channel, multi-template event reconstruction for SuperCDMS data using machine learning

SuperCDMS SNOLAB uses kilogram-scale germanium and silicon detectors to search for dark matter. Each detector has Transition Edge Sensors (TESs) patterned on the top and bottom faces of a large crystal substrate, with the TESs electrically grouped into six phonon readout channels per face. Noise correlations are expected among a detector's readout channels, in part because the channels and their readout electronics are located in close proximity to one another. Moreover, owing to the large size of the detectors, energy deposits can produce vastly different phonon propagation patterns depending on their location in the substrate, resulting in a strong position dependence in the readout-channel pulse shapes. Both of these effects can degrade the energy resolution and consequently diminish the dark matter search sensitivity of the experiment if not accounted for properly. We present a new algorithm for pulse reconstruction, mathematically formulated to take into account correlated noise and pulse shape variations. This new algorithm fits N readout channels with a superposition of M pulse templates simultaneously - hence termed the N$\times$M filter. We describe a method to derive the pulse templates using principal component analysis (PCA) and to extract energy and position information using a gradient boosted decision tree (GBDT). We show that these new N$\times$M and GBDT analysis tools can reduce the impact from correlated noise sources while improving the reconstructed energy resolution for simulated mono-energetic events by more than a factor of three and for the 71Ge K-shell electron-capture peak recoils measured in a previous version of SuperCDMS called CDMSlite to $<$ 50 eV from the previously published value of $\sim$100 eV. These results lay the groundwork for position reconstruction in SuperCDMS with the N$\times$M outputs.

Albakry, M. F. [British Columbia U.; TRIUMF]↗

Solid-State Mixed-Potential Electrochemical Sensors for Natural Gas Leak Detection and Quality Control (Final Technical Report)

Mitigation of methane emissions are a critical factor to limiting the impact of the natural gas industry on global climate change. Throughout the period of 2020-2024, the University of New Mexico and its commercialization partner and subcontractor, SensorComm Technologies, Inc. (SCT), have worked together to develop a low-cost Artificial Intelligence (AI)-driven Internet of Things (IoT)-based multi-gas sensor platform for methane emissions detection. In the final year of the project, we extended this work to include hydrogen detection in support of a transition to a hydrogen economy where hydrogen could be transported through existing natural gas infrastructure. Mixed potential electrochemical sensors were first prototyped by ceramic additive manufacturing and then transitioned to conventional ceramic manufacturing tape casting and screen-printing technologies in preparation for mass production. Demonstrated limits of detection of 5 ppm of methane in natural gas and 1 ppm of hydrogen were measured. These limits of detection are among the lowest of solid-state electrochemical sensors that have been reported in the literature or available in the industry. Machine learning algorithms were developed to identify natural gas mixtures with > 98% accuracy level and quantify methane concentrations at 97% accuracy. The presence of hydrogen could also be identified, and its concentration quantified at these accuracy levels. These algorithms were optimized for running on portable computing hardware which enabled > 1 Hz processing rates. A portable packaged IoT system was integrated with the electrochemical sensor in collaboration with SCT. The package consists of readout electronics with < 1 mV resolution, sensor temperature control, and data transmission over cellular wireless and/or Wi-Fi networks. Field testing was performed in two rounds at Colorado State University’s Methane Emissions Technology Evaluation Center (CSU METEC). The first round of testing demonstrated successful measurements of methane from an underground natural gas leak of 20 standard liters per minute (SLPM), which agreed with previously published literature using more sophisticated and expensive analytical equipment. The second round of testing showed that an above ground leak of 2 SLPM of hydrogen could be detected at 32 ft. This project has resulted in six published peer reviewed journal articles, over ten presentations at professional conferences, and one full patent application filed in 2023. Future work on this project includes increased sensitivity, higher production yields, and applications in the hydrogen safety and flare emissions monitoring spaces.

03 NATURAL GAS↗