Search NASA⌕ Search

SEARCH · Search NASA

Results for “sampling algorithm”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 379 records · Page 21

Wasserstein Normalized Autoencoder for Anomaly Detection in ProtoDUNE Vertical-Drift Detector

ProtoDUNE Vertical Drift needs a selective triggering algorithm. The detector sits on Earth's surface, so cosmic activity dominates its data. Our goal in this paper is to trigger on neutrino events more robustly than the current deployed Analog-to-Digital Converter Simple Window (ADCSW) model and, eventually, search for signals of Beyond Standard Model (BSM) physics at DUNE as our ultimate North Star objective. As a step towards this goal, we evaluate a Wasserstein Normalized Autoencoder (WNAE) on simulated collection-plane only windows of shape $1\times10\times10$ where Neutrinos act as our BSM-proxy and Cosmic-ray Muons serve as our learned background. The network parameters are fitted using only cosmic-ray muon events as background in order to maintain an unsupervised pipeline. Training uses finite-step Langevin $x^-$ samples, positive-sample reconstruction energy, and an empirical sliced $2$-Wasserstein objective to learn a normalized Boltzmann energy model. We then calibrate on a nominal $5\,\mathrm{Hz}$ operating threshold calculated from cosmic validation data. Both WNAE and ADCSW accept 311 of 194,083 held-out cosmic background events at this $5\,\mathrm{Hz}$ threshold. We found that WNAE accepts 9,677 of 34,634 neutrino-proxy events $(27.9\pm0.24)\%$, compared with 10,076 $(29.1\pm0.24)\%$ for ADCSW, an observed WNAE-minus-ADCSW difference of $-1.15\%$. At another nominal $2\,\mathrm{Hz}$ target threshold, the corresponding efficiencies are $(20.5\pm0.22)\%$ and $(22.6\pm0.22)\%$, respectively. Of the WNAE-selected neutrino proxies at $5\,\mathrm{Hz}$, $(20.8\pm0.4)\%$ of the classified neutrino-proxy events are unique to WNAE, where the uncertainty is an absolute binomial standard error of $0.4\%$.

Zheng, Jake [U. Chicago (main)] (ORCID:00090002189↗

Wasserstein Normalized Autoencoder for Anomaly Detection in ProtoDUNE Vertical-Drift Detector

ProtoDUNE Vertical Drift needs a selective triggering algorithm. The detector sits on Earth's surface, so cosmic activity dominates its data. Our goal in this paper is to trigger on neutrino events more robustly than the current deployed Analog-to-Digital Converter Simple Window (ADCSW) model and, eventually, search for signals of Beyond Standard Model (BSM) physics at DUNE as our ultimate North Star objective. As a step towards this goal, we evaluate a Wasserstein Normalized Autoencoder (WNAE) on simulated collection-plane only windows of shape $1\times10\times10$ where Neutrinos act as our BSM-proxy and Cosmic-ray Muons serve as our learned background. The network parameters are fitted using only cosmic-ray muon events as background in order to maintain an unsupervised pipeline. Training uses finite-step Langevin $x^-$ samples, positive-sample reconstruction energy, and an empirical sliced $2$-Wasserstein objective to learn a normalized Boltzmann energy model. We then calibrate on a nominal $5\,\mathrm{Hz}$ operating threshold calculated from cosmic validation data. Both WNAE and ADCSW accept 311 of 194,083 held-out cosmic background events at this $5\,\mathrm{Hz}$ threshold. We found that WNAE accepts 9,677 of 34,634 neutrino-proxy events $(27.9\pm0.24)\%$, compared with 10,076 $(29.1\pm0.24)\%$ for ADCSW, an observed WNAE-minus-ADCSW difference of $-1.15\%$. At another nominal $2\,\mathrm{Hz}$ target threshold, the corresponding efficiencies are $(20.5\pm0.22)\%$ and $(22.6\pm0.22)\%$, respectively. Of the WNAE-selected neutrino proxies at $5\,\mathrm{Hz}$, $(20.8\pm0.4)\%$ of the classified neutrino-proxy events are unique to WNAE, where the uncertainty is an absolute binomial standard error of $0.4\%$.

Zheng, Jake [Chicago U.] (ORCID:0009000218901379)↗

Velocity reconstruction in the era of DESI and Rubin/LSST. II. Realistic samples on the light cone

Reconstructing the galaxy peculiar velocity field from the distribution of large-scale structure plays an important role in cosmology. On one hand, it gives us an insight into structure formation and gravity; on the other, it allows us to selectively extract the kinetic Sunyaev-Zel’dovich (kSZ) effect from cosmic microwave background maps. In this work, we employ high-accuracy synthetic galaxy catalogs on the light cone to investigate how well we can recover the velocity field when utilizing the three-dimensional spatial distribution of the galaxies in a modern large-scale structure experiment such as the Dark Energy Spectroscopic Instrument (DESI) and the Rubin Observatory Legacy Survey of Space and Time. In particular, we adopt the standard technique used in baryon acoustic oscillation analysis for reconstructing the Zel’dovich displacements of galaxies through the continuity equation, which yields a first-order approximation to their large-scale velocities. We investigate variations in the number density, bias, mask, area, redshift noise, and survey depth, as well as modifications to the settings of the standard reconstruction algorithm. Since our main goal is to provide guidance for planned kSZ analysis between DESI and the Atacama Cosmology Telescope, we apply velocity reconstruction to a faithful representation of DESI spectroscopic and photometric targets. We report the cross-correlation coefficient between the reconstructed and the true velocities along the line of sight. For the DESI Y1 spectroscopic survey, we expect the correlation coefficient to be r ≈ 0.64, while for a photometric survey with δ z /(1+z) = 0.02, as is approximately the case for the Legacy Survey used in the target selection of DESI galaxies, r shrinks by half to r ≈ 0.31. Here, we hope the results in this paper can be used to inform future kSZ stacking studies and other velocity reconstruction analyses planned with the next generation of cosmology experiments.

79 ASTRONOMY AND ASTROPHYSICS↗

Effective optimization of atomic decoration in giant and superstructurally ordered crystals with machine learning

Crystals with complicated geometry are often observed with mixed chemical occupancy among Wyckoff sites, presenting a unique challenge for accurate atomic modeling. Similar systems possessing exact occupancy on all the sites can exhibit superstructural ordering, dramatically inflating the unit cell size. In this work, a crystal graph convolutional neural network (CGCNN) is used to predict optimal atomic decorations on fixed crystalline geometries. This is achieved with a site permutation search (SPS) optimization algorithm based on Monte Carlo moves combined with simulated annealing and basin-hopping techniques. Our approach relies on the evidence that, for a given chemical composition, a CGCNN estimates the correct energetic ordering of different atomic decorations, as predicted by electronic structure calculations. This provides a suitable energy landscape that can be optimized according to site occupation, allowing the prediction of chemical decoration in crystals exhibiting mixed or disordered occupancy, or superstructural ordering. Verification of the procedure is carried out on several known compounds, including the superstructurally ordered clathrate compound Rb8Ga27Sb16 and vacancy-ordered perovskite Cs2SnI6, neither of which was previously seen during the neural network training. In addition, the critical temperature of an order–disorder phase transition in solid solution CuZn is probed with our SPS routines by sampling site configuration trajectories in the canonical ensemble. This strategy provides an accurate method for determining favorable decoration in complex crystals and analyzing site occupation at unprecedented speed and scale.

Chemistry↗

Development of High-Granularity Dual-Readout Calorimetry with psec Timing

Dual-readout and particle flow algorithm (PFA) are technologies proposed for precise jet energy measurement in future colliders. While PFA requires highly granular calorimeters, dual-readout has mainly been used with fiber-based calorimeters that do not have highly segmented capabilities. It is still non-trivial to combine these two technologies in one calorimeter system because of the use of fibers in most of the dualreadout calorimeters, which is not compatible with the high granularity requirement of PFA technologies. The aim of this study is to develop a novel calorimetry that combines dual-readout and PFA by adopting a highly segmented tile-based configuration. This paper compares the improvement of energy resolution using dual-readout approach across several configurations of highly granular hadron calorimeters through simulation. Results indicate that setups with fine sampling and close placement of scintillators and Cherenkov detectors improve dual-readout performance.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

A Behavioral Robotics Approach to Radiation Mapping Using Adaptive Sampling

Radiation mapping is a desirable task to automate because of the inherent risks involved and its tedious nature. A novel system was designed to address this by combining various existing technologies, utilizing behavior-based robotics and Bayesian optimization. The system uses a quadruped robot equipped with a manipulator and gamma detector to take measurements at locations that are selected based on the uncertainty of a surrogate model used to estimate the true radiation field. The robot uses input from the world with depth cameras to avoid collisions with the robot’s body, and unreachable points for the end effector are addressed by both allowing for a soft collision with the environment to occur, prompting the system to abandon that point, and varying the exploration tendency of the optimization based on consecutive collisions. This approach provides unique traversability and adaptability over other strategies in the literature. Experiments were performed by placing a Cesium-137 source on the ground and varying geometric setups and an optimization parameter demonstrating the adaptability to diverse environments and the increased robustness resulting from the designed behavior. The results additionally demonstrate that dynamically adjusting the optimization algorithm’s exploration tendency based on the arm’s collision history improves the system’s ability to navigate cluttered environments and construct accurate radiation maps without getting stuck in unreachable areas.

Adams, Joel↗

Modeling Approach for the Aluminum-clad Dry Storage Pilot using HFIR Fuel

To confirm that the dry storage of aluminum-clad research reactor spent nuclear fuel (ASNF) will remain within the safety envelope after applied drying schemes and that the resulting evolution of the gas space composition, temperature, and pressure conditions are understood, a dry storage pilot project is being established. The pilot will incorporate an instrumented lid for discrete interval or for on-demand gas composition and temperature monitoring of two DOE Standard Canisters (DSCs) loaded with three High Flux Isotope Reactor (HFIR) inner cores per DSC. Each DSC would be subjected to a separate alternative candidate drying scheme. Canisters will undergo 1 to 5 years of monitoring, including internal temperature and gas sampling to track pressure and composition changes. This report outlines the approach for modeling the ASNF-in-canister behavior in terms of evolving gas space conditions for the ASNF dry storage pilot using HFIR fuel. The ASNF has an adherent surface oxyhydroxide layer comprised of boehmite/bayerite that generates hydrogen when subjected to irradiation. Three-dimensional multi-physics computational fluid dynamics simulations will be executed to compute the thermal field within the DSC and provide inputs to a chemical model employed to compute pressure buildup as hydrogen is generated in the system. Implemented in Cantera, the chemical model solves gas phase and aluminum oxyhydroxide surface-mediated radiolysis reactions. Gas phase reactions are sourced from Wittman and Hanson (2015), whereas surface-mediated reactions are incorporated by fitting experimental data using an optimization algorithm (Abboud, 2023). Water radiolysis reactions from Wren and Ball (2001) are adopted with modifications as described in Abboud (2023c). Understanding the effect of the hydrogen buildup over time is important for long-term storage safety considerations. Modeling results will include the canister pressure, temperature, and composition evolution from the initial helium backfill with the addition of radiolytically-evolved chemical species (e.g., hydrogen and oxygen). The specific HFIR cores for the pilot program have not yet been selected, and the overall design is still in development. The CFD-chemical model used for this work will be based on prior models with necessary updates to allow for improved accuracy and efficiency. The experimental data obtained from the HFIR demonstration will be used to improve and validate the computational models to predict the ASNF-in-canister behavior.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

Rapid Bayesian High Entropy Alloy Designs Fabricated via Wire Arc Additive Manufacturing

Purpose: This project seeks to demonstrate a new high-throughput (rapid) alloy design technique applied to creating new high entropy alloys (HEAs) for extreme environments. High entropy alloys shift the design paradigm from being focused on a single principal element (e.g. nickel-based alloys) to target alloys that include high atomic fractions (X >10%) of multiple elements. These HEA materials can exhibit sluggish diffusion and enhanced corrosion resistance, ideal for potential applications in advanced ultra supercritical (A-USC) steam cycles for power generation. Scope: The addition of multiple elements in high atomic fractions creates an enormous design space that cannot easily be investigated by traditional material design strategies such as designed of experiments (DOE). This project utilizes a Bayesian machine learning algorithm that has been modified to work with calculation of phase diagrams (CALPHAD) software. This Bayesian algorithm reduces manual inputs and increase the likelihood of achieving an optimal solution. Compositional inputs to this algorithm will be assessed using existing material property models for high temperature strength and corrosion resistance. The target for alloy performance will be a 15% (~100 ⁰C) increase in allowable service temperature beyond heat-resistant stainless steels while maintaining or improving alloy cost and corrosion resistance. Haynes 230 was selected as a baseline, which is 57 wt% Ni with 22 wt% Cr 14 wt% W, and 2 wt% Mo as solid solution strengtheners. In addition to rapid design via Bayesian machine learning, the alloys were rapidly fabricated using a multi-wire arc additive manufacturing (mWAAM) technique which allows for precise control of alloy composition and assessing of alloy design “windows” to study composition effects. Build speeds for wire-arc additive processes are among the highest for additive technologies enabling rapid and reliable sample fabrication when compared to conventional methods such as arc button melting. The mWAAM samples will be rapidly characterized via instrumented indentation for room temperature modulus and strength and for elevated temperature strength via hot hardness tests. After being screened with hardness testing, potential alloys will be further evaluated with conventional microscopy techniques including scanning electron microscopy (SEM) and transmission electron microscopy (TEM) to assess agreement with modeling results. The most promising compositions will also be evaluated by printing full sized tensile specimens for mechanical behavior tests at elevated temperatures. Results: Bayesian machine learning of a single performance function was initially used to optimize five performance metrics: 1) single phase stability, 2) yield strength, 3) creep resistance (low diffusion coefficient), 4) freezing range (weldability), and 5) material cost. The single performance function was suboptimal as assumptions had to be made about the results while formulating the optimization. A goal-oriented Bayesian optimization strategy (Hanaoka, 2021) was implemented with CALPHAD for use with the five metrics above. This multi-objective Bayesian optimization (MOBO) enabled the design of NiCrCoFe alloys with V and W additions. A base composition of NiCoCr was selected as Ni provides a stable FCC matrix, Cr aids corrosion/oxidation resistance, and Co is a solid-solutions strengthener that also improves creep by increasing the activation energy. Fe helps reduce diffusion coefficients and cost. Finally, V and W were selected for their reasonable solubility and high atomic misfit to aid in solid solution strengthening. Cracking of the mWAAM specimens was an early issue, and the Easton solidification cracking model (Easton et al., 2014a) was selected for addition to the MOBO function. High performing alloys fabricated by mWAAM included Ni 28 Cr 25 Co 26 Fe 15 V 8 and Ni 62 Cr 18 Co 1 Fe 3 W 15 . It was observed that even after adapting the mWAAM process for W, the W did not fully dissolve. To fully evaluate the Ni 62 Cr 18 Co 1 Fe 3 W 15 composition, a cored wire (80-20 NiCr sheath/powder core) was manufactured and printed via WAAM, and HIP’ing was utilized to homogenize and densify the printed alloy. The V and W alloys produced met metrics 1 (solid solution), 4 (solidification cracking), and 5 (cost). However, an unmodeled mechanism of thermal stress cracking was identified in the WAAM produced materials, perhaps exacerbated by the lack of grain boundary strengthening elements (B, C). Conclusions & Recommendations: A high-throughput (rapid) alloy design technique was applied to designing and manufacturing new high entropy alloys (HEAs) for extreme environments utilizing MOBO and mWAAM. The developed process was rapid and effective in addressing the mechanisms included in the model. The lack of grain boundary strengthening element additions (e.g., B, C) was a simplification that likely produced thermal stress cracking that turned into a large part of the investigation. Additions on the order of 0.005 wt% B and 0.05 wt% C likely would have minimized thermal stress grain boundary cracking. Overall, the high throughput design strategy is promising for rapid design of metrics-driven alloys for advanced ultra supercritical (A-USC) steam cycles for power generation. The MOBO and mWAAM process could be commercialized to accelerate metrics-driven alloy design. In addition, the cored-wire process utilized for scale-up is a promising high-volume process for WAAM alloy development and scale-up.

36 MATERIALS SCIENCE↗

RU Net for Automatic Characterization of TRISO Fuel Cross Sections

TRistructural ISOtropic (TRISO) particle fuel is a type of nuclear fuel known for its high-temperature and high-burnup performance. Each sub-millimeter diameter TRISO particle consists of uranium-oxycarbide (UCO) or UO2 fuel kernel, coated with buffer, inner pyrolytic carbon (IPyC), silicon carbide (SiC), and outer pyrolytic carbon (OPyC) layers. The SiC layer acts as the main containment barrier for the TRISO particle to retain the fission products, while the IPyC and OPyC layers provide additional barriers to the release of fission products, especially fission gases. During irradiation, phenomena like kernel swelling, buffer densification, and IPyC fracture may impact fuel performance. Post-irradiation microscopy on entire compact cross sections or samples of individual particles deconsolidated from compacts is often used to identify these irradiation-induced changes in morphology. However, each fuel compact generally contains thousands of TRISO particles. To get statistical information on these phenomena, it is cumbersome work if done manually. For example, to get information about swelling/densification behaviors of different layers or kernels after irradiation, researchers previously manually measured the perimeter of each TRISO layer in hundreds of particles after four rounds of iterative grinding and polishing encompassing more than 2000 cross-section images for a total of four fuel compacts. To attempt to reduce the subjectivity inherent in that process and accelerate data analysis, we conducted a study on the automatic TRISO layer segmentation on cross-sectional microscopic images using Convolutional Neural Networks (CNNs). CNNs are a class of machine learning algorithms specifically designed for processing structured grid data that have gained popularity in recent years due to their remarkable performance in various computer vision tasks, including image classification, object detection, and image segmentation. In this research, we have generated the large irradiated TRISO layer dataset with more than 2000 cross-section TRISO microscopic images and the corresponding annotated images. Based on these annotated images, we have employed different CNNs for automatic segmentation of different TRISO layers. These include RU-Net (developed in this study), as well as three existing architectures: U-Net, Residual Network (ResNet), and Attention U-Net. The preliminary results show that the model based on RU-Net has the best performance in terms of intersection-over-union (IoU). Through the aid of these CNN models, we can expedite the analysis of TRISO particle cross-sections, significantly reducing the manual labor involved and improving the objectivity of the segmentation results.

Convolutional Neural Networks↗

Data Assimilation for Robust UQ Within Agent-Based Simulation on HPC Systems

Agent-based simulation provides a powerful tool for in silico system modeling. However, these simulations do not provide built-in methods for uncertainty quantification (UQ). Within these types of models a typical approach to UQ is to run multiple realizations of the model then compute aggregate statistics. This approach is limited due to the compute time required for a solution. When faced with an emerging biothreat, public health decisions need to be made quickly and solutions for integrating near real-time data with analytic tools are needed. We propose an integrated Bayesian UQ framework for agent-based models based on sequential Monte Carlo sampling. Given streaming or static data about the evolution of an emerging pathogen this Bayesian framework provides a distribution over the parameters governing the spread of a disease through a population. These estimates of the spread of a disease may be provided to public health agencies seeking to abate the spread. By coupling agent-based simulations with Bayesian modeling in a data assimilation, our proposed framework provides a powerful tool for modeling dynamical systems in silico. We propose a method which reduces model error and provides a range of realistic possible outcomes. Moreover, our method addresses two primary limitations of ABMs: the lack of UQ and an inability to assimilate data. Our proposed framework combines the flexibility of an agent-based model with UQ provided by the Bayesian paradigm in a workflow which scales well to HPC systems. We provide algorithmic details and results on a simulated outbreak with both static and streaming data.

Spannaus, Adam [ORNL] (ORCID:0000000225213657)↗

From pixels to patterns: Coupling Optical Coherence Tomography and machine learning for monitoring coastal wetland root systems

Coastal wetlands are crucial in shoreline stabilization, carbon sequestration, and storm protection. Yet, due to limitations in traditional destructive sampling techniques, the belowground biomass (live root mass) and necromass (dead and decaying roots) remain difficult to assess in coastal wetlands, limiting our understanding on coastal resilience, nutrient cycling, and soil structure. This study employs Optical Coherence Tomography (OCT) as a high-resolution imaging technique to analyze root biomass and necromass in the Terrebonne Basin, Louisiana. A Random Forest (RF) model was developed to classify root health states based on OCT-derived features, achieving an accuracy of 70% in distinguishing live from dead root segments. The results demonstrate that OCT, combined with ML, offers a promising novel approach to root analysis, providing fine-scale insights into root morphology and decay patterns that are not easily captured by conventional methods. This research lays the foundation for future integration of OCT with complementary imaging modalities such as X-ray Computed Tomography (XCT) and advanced ML algorithms to enhance classification accuracy and scalability. Future work aims to expand the dataset diversity across different wetland types and apply the methodology for large-scale, repeatable assessments of root biomass turnover and accumulation, with important implications for wetland monitoring, conservation, and restoration under changing environmental conditions.

AI/ML↗

Fidelity-preserving enhancement of ptychography with foundational text-to-image models

Ptychographic phase retrieval enables high-resolution imaging of complex samples but often suffers from artifacts such as grid pathology and multislice crosstalk, which degrade reconstructed images. We propose a plug-and-play (PnP) framework that integrates physics model-based phase retrieval with text-guided image editing using foundational diffusion models. By employing the alternating direction method of multipliers, our approach ensures consensus between data fidelity and artifact removal subproblems, maintaining physical consistency while enhancing image quality. Artifact removal is achieved using a text-guided diffusion image editing method (LEDITS++) with a pre-trained foundational diffusion model, allowing users to specify artifacts for removal in natural language. Demonstrations on simulated and experimental datasets show significant improvements in artifact suppression and structural fidelity, validated by metrics such as peak signal-to-noise ratio and diffraction pattern consistency. This work highlights the combination of text-guided generative models and model-based phase retrieval algorithms as a transferable and fidelity-preserving method for high-quality diffraction imaging.

image editing↗

Implementation of compound refractive lenses for large field-of-view x-ray phase-contrast imaging during hypervelocity impact experiments

Synchrotron x-ray phase-contrast imaging (XPCI) offers time-resolved visualization of dynamic compression phenomena, but its intrinsically small field-of-view (FOV) limits the time that key features remain in frame. A novel approach to enlarge the FOV is achieved by positioning a two-dimensional parabolic compound refractive lens (CRL) upstream of the sample to deliberately defocus the white beam. Ray-tracing simulations and XPCI measurements show that this CRL configuration can expand the beam by ∼50% vertically and ∼15% horizontally based on the full width at half-maximum of the beam. Implementing the CRL, however, attenuates the photon flux and lowers signal-to-noise ratio (SNR). Task-based analysis using a calibration grid (30 μm dots) showed that both setups fail to consistently meet the Rose criterion (SNR ≥ 5) for features of this size in single-bunch imaging. Extrapolating the measured SNR Rose values suggests that the minimum consistently detectable feature lies closer to 30–40 μm for the standard XPCI setup and above 40 μm for CRL-XPCI. Despite this limitation, the CRL configuration nearly doubles the illuminated area, enabling simultaneous tracking of front and rear observations of boron carbide targets subjected to rod and sphere impacts at 1.0–2.6 km/s. Image tracking algorithms and photonic Doppler velocimetry were used to measure penetration and rear-surface velocity histories. Together, these measurements capture crack fronts, penetration, and material breakout, offering new benchmark data for validating high-strain-rate constitutive models of ceramic materials.

Ceramic materials↗

Kekulé valence bond order in the honeycomb lattice optical Su-Schrieffer-Heeger model and its relevance to graphene

We perform sign-problem-free determinant quantum Monte Carlo simulations of the optical Su- Schrieffer-Heeger model on a half-filled honeycomb lattice. In particular, we investigate the model’s semi-metal (SM) to Kekulé Valence Bond Solid (KVBS) phase transition at zero and finite temper- atures as a function of phonon energy and interaction strength. Using hybrid Monte Carlo sampling methods we can simulate the model near the adiabatic regime, allowing us to access regions of parameter space relevant to graphene. Our simulations suggest that the SM-KVBS transition is weakly first-order at all temperatures, with graphene situated close to the phase boundary in the SM region of the phase diagram. Furthermore, our results highlight the important role bond-stretching phonon modes play in the formation of KVBS order in strained graphene-derived systems.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Jet fragmentation function and groomed substructure of bottom quark jets in proton-proton collisions at 5.02 TeV

A measurement of the substructure of bottom quark jets (b jets) in proton-proton (pp) collisions is presented. The measurement uses data collected in pp collisions at $\sqrt{s}=5.02$ TeV, with a low number of simultaneous interactions per bunch crossing, recorded by the CMS experiment in 2017, corresponding to an integrated luminosity of 301 pb −1 . An algorithm to identify and cluster the charged decay daughters of b hadrons is developed for this analysis, which facilitates the exposure of the gluon radiation pattern of b jets using iterative Cambridge-Aachen declustering. The soft-drop-groomed jet radius, R g , and momentum balance, z g , of b quark jets are presented. These observables can be used to test perturbative quantum chromodynamics predictions that account for mass effects. Because the b hadron is partially reconstructed from its charged decay daughters, only charged particles are used for the jet substructure studies. In addition, a jet fragmentation function, z b,ch , is measured, which is defined as the distribution of the ratio of the transverse momentum (p T ) of the partially reconstructed b hadron with respect to the charged-particle component of the jet p T . The substructure variable distributions are unfolded to the charged-particle level. The b jet substructure is compared to the substructure of jets in an inclusive jet sample that is dominated by light-quark and gluon jets in order to assess the role of the b quark mass. A strong suppression of emissions at small R g values is observed for b jets when compared to inclusive jets, consistent with the dead-cone effect. The measurement is also compared with theoretical predictions from Monte Carlo event generators. This is the first substructure measurement of b jets that clusters together the b hadron decay daughters independent of the b hadron species and decay channel.

boosted jets↗

Multiobjective Constrained Symbolic Regression for Predictive Modeling of Material Creep Behavior

When creep testing is repeated on samples of the same alloy under the same parametric conditions (i.e., stress and temperature), the resulting strain/time curves can vary from each other considerably as shown in Figure 1 [1]. The time required to creep test a material to rupture can extend to the order of years. Because of this, a numerical model that can quickly analyze the incomplete results of an ongoing experiment to predict 1) the incomplete portion of the strain/time curve leading up to the rupture point and 2) the rupture point itself would be of great utility to the materials community. Such a model has the potential to save 1) the time required to finish running the experiment to rupture 2) the associated monetary cost of finishing said experiment. Furthermore, it would be advantageous if the predictive model could give a parametric function modeling strain/time curves for material scientists to investigate the impact of the temperature and stress parameters on the resulting creep behavior. This work introduces a piecewise symbolic regression algorithm to predict the remainder of the strain/time curve. Preliminary results show good model performance.

36 MATERIALS SCIENCE↗

Hierarchical Bayesian Modeling for Cosmology: Can NPE reliably replace MCMC?

Hierarchical neural posterior estimation has its place Hierarchical Bayesian Modeling (HBM) combined with MCMC algorithms has been shown to provide more robust and accurate inference for real-world phenomena in which nature takes a nested form. However, MCMC-based inference can be computationally expensive, and its performance often suffers for complex posterior geometries. These costs are especially pertinent for HBM. Studies have recently demonstrated the potential for a flexible, expressive, and amortized hierarchical neural posterior estimator (HNPE) built on Normalizing Flows. These studies have mostly been performed on simple datasets, or they focus on a single parameter from each level of the hierarchy. A systematic study analyzing how both hierarchical methods compare for more complex and realistic datasets is necessary before applying HNPE for scientific measurements. Here, we re-explore the theory behind HNPE and conduct comparative numerical experiments of HNPE and MCMC-based HBM methods on real and synthetic data, including strong gravitational lensing simulations. In particular, we use a suite of diagnostics to show trade-offs in terms of accuracy, precision, time to train or sample, reproducibility, and the need for expert domain knowledge. Especially for higher dimensional and complex posteriors, HNPE is expected to drastically improve on time for inference, accuracy, and precision with an upfront training time cost.

Hur, Rachel [Chicago U.] (ORCID:000900089890445X)↗

Synthesis of ARM User Facility Surface Rainfall Datasets to Construct a Best Estimate Value Added Product (PrecipBE)

Surface precipitation measurements are essential for Earth system model (ESM) evaluation and understanding cloud processes. An ever-growing need for robust, temporally evolving, and easy-to-use statistical datasets provides motivation for a baseline ground-based precipitation properties data product. The U.S. Department of Energy Atmospheric Radiation Measurement (ARM) user facility operates an extensive suite of precipitation instruments with various sensitivities and operating mechanisms, which render the decision of which instrument to use based on one or more fixed thresholds challenging and prone to errors and bias. Using a long-term instrument inter-comparison from a unique per-precipitation event perspective, rather than instantaneous sample comparison, we demonstrate that ARM rainfall-measuring instruments are generally consistent with each other at the statistical level. Inter-instrument deviations at the single event level can be large, especially for specific rainfall event properties such as maximum precipitation rates. A machine-learning (ML) analysis using a random forest regressor indicates that in some cases, depending on instrument, local site climatology, and/or specific deployment configuration, certain atmospheric state variables influence the measured quantities in an unpredictable manner. Thus, a-priori weighting of different instruments does not necessarily lead to more accurate and less biased synthesis of instrument data. These results motivate the design of the ARM precipitation best-estimate (PrecipBE) value-added product, which incorporates all valid precipitation data while considering data quality and other instrument limitations. PrecipBE consists of time series and tabular statistics datasets in an easy-to-use and insightful per-precipitation event format. It provides a large set of precipitation event properties supplemented with ancillary data from ARM datasets that correspond to the detected precipitation events. We describe the PrecipBE algorithm and demonstrate its use via the examination of a single-day output as well as a long-term trend analysis of precipitation events at the ARM Southern Great Plains (SGP) site, covering more than 30 years of data. The trend analysis tentatively suggests a long-term temporal tendency for mainly shorter and less intense precipitation events at the SGP site, but a long-term increase in annual rainfall by more than 36 mm (5 %) per decade. This rainfall trend is catalyzed primarily by more extreme event properties of relatively rare, intense precipitation events, with event total and 1 min maximum precipitation rate at a 1 year timeframe increasing up to 5 mm and 9 mm h −1 (several percent) per decade, respectively. While the currently available PrecipBE datasets (at https://adc.arm.gov/discovery/, last access: 8 December 2025) cover rainfall from multiple ARM deployments up to March 2025, PrecipBE is planned to be expanded to include solid-phase precipitation and will soon become an operational product with a several-day lag from real-time. We invite the ARM user community to leverage this new product and welcome user feedback to further enhance the dataset.

Silber, Israel [Pacific Northwest National Laborat↗