Search NASA⌕ Search

SEARCH · Search NASA

Results for “Reasoning”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

Probabilistic inference in very large universes

Our current favored cosmological theories allow for the striking and controversial possibility that the observable universe is just a small part of a much larger universe in which parameters that describe the effective, low-energy laws of physics vary from one region to another. The controversy is largely driven by the fact that such a “very large universe” is mostly observationally inaccessible to us, so the issue arises of how we can reasonably assess a theory that describes such a universe. In this paper, we propose a Bayesian method for theory assessment based on theory-generated probability distributions for our observations. We focus on the principles that define this method, leaving aside concerns about how, in practice, one would carry out the required calculations. (One important issue that we set aside is the measure problem.) We argue that cosmological theories can be tested by the standard method of Bayesian updating, but we need to use theoretical predictions for “first-person” probabilities—that is, probabilities that we should use for our observations, taking into account all relevant selection effects. These selection effects can vary from one observer to another and can vary with time, so, in principle, first-person probabilities are defined for each observer instant—an observer at a specific instant of time. Calculations of first-person probabilities should take into account everything that the observer believes about herself and her surroundings, which we refer to as her subjective state. If the universe is very large, a theory might predict that there are many observer instants in the same subjective state; we argue that first-person probabilities should be calculated using a principle of self-locating indifference (PSLI), the assumption that any real observer should make predictions for her future as if she were chosen randomly and uniformly from the theoretically predicted observer instants that share her subjective state. We believe the PSLI is intuitively very reasonable, but we also argue that, if the theory is correct, the use of this principle maximizes the expected fraction of observers who will make correct predictions. A further complication is that cosmological theories are not expected to fully predict the detailed properties of the universe, but rather will predict a set of possible universes, each with a probability. Different possible universes will generically have different numbers of observers. We argue that, in the calculation of first-person probabilities, the probability for each possible universe should be weighted by the number of observer instants in the specified subjective state that it contains. These issues have been controversial in the literature, so we also provide a rebuttal to the claim that principles like the PSLI involve a “selection fallacy”; a rebuttal to what we dub the principle of required certainty; an argument rejecting theories that predict a preponderance of Boltzmann brains; a rebuttal to a parable about humans and Jovians used by Hartle and Srednicki to argue that assumptions of typicality can lead to absurd consequences; and, finally, a discussion about how the use of “old evidence” can be fit into a Bayesian mold.

Azhar, Feraz [University of Notre Dame, IN (United↗

Primordial black hole dark matter: A quantitative parameter sensitivity comparison across formation mechanisms and particle candidates

Primordial black holes (PBHs) in the asteroid-mass window ( 10 17 – 10 22 g ) can account for all of the dark matter without violating any observational constraint, yet are routinely dismissed as fine-tuned. I put that dismissal to the test by applying three complementary sensitivity measures uniformly across a broad landscape: three noninflationary PBH production mechanisms, six classes of inflationary PBH models, and seven particle dark matter benchmarks, all evaluated against the same observable target. Three distinct naturalness universality classes emerge, determined entirely by the analytic structure of the abundance map rather than by the nature of the dark matter candidate. Biased-domain-wall PBHs, in their least model-dependent (free- V b ) form, have the same low sensitivity, Δ = 4.5 , as off-resonance weakly interacting massive particles and freeze-in particles ( Δ = 2 ), a sensitivity that, because it is constant over the entire parameter space of the construction, also coincides trivially with its own Wilson-normalized average within that parameter space (Section Definition and conventions), an equivalence that concerns only the space over which Δ is computed and is not a naturalness statement about the construction as a whole; a further reduction to Δ = 2 is possible only under the additional, independently motivated but not required, assumption that the domain-wall bias is generated by Planck-suppressed operators; early matter-domination PBHs occupy an intermediate tier alongside coannihilating weakly interacting massive particles (WIMPs), unified by a structural identity in which the sensitivity measure equals the logarithm of the ratio of the formation scale to the matter–radiation equality scale; first-order phase transition PBHs, once the more accurate super-exponential collapse probability is used in place of the single-exponential approximation, instead belong to the same highly sensitive tier as resonant WIMP annihilation and single-field inflationary collapse, for a structurally distinct reason; single-field ultraslow-roll inflationary collapse is severely tuned for a distinct reason: a double exponential in which the power spectrum amplitude is itself exponentially sensitive to the inflaton potential coefficients, on top of the exponential collapse sensitivity of the abundance map. My main conclusion is that the claim that PBH dark matter is generically fine-tuned conflates the worst case with a landscape spanning every naturalness tier. The Barbieri-Giudice sensitivity computed here and the Wilson-normalized measure of Iovino and Riotto answer distinct and mutually consistent questions about the same construction, a distinction I clarify within the two-layer decomposition.

Profumo, Stefano [University of California, Santa ↗

CMB lensing and Ly⁢ α forest cross bispectrum from DESI’s first-year quasar sample

The squeezed cross-bispectrum B κ,Ly α between the gravitational lensing in the cosmic microwave background and the 1D Ly α forest power spectrum can constrain bias parameters and break degeneracies between σ8 and other cosmological parameters. We detect B κ,Ly ⁢α with 4.8⁢σ significance at an effective redshift z eff =2.4 using Planck PR3 lensing map and over 280,000 quasar spectra from the Dark Energy Spectroscopic Instrument’s first-year data. We test our measurement against metal contamination and foregrounds such as Galactic extinction and clusters of galaxies by deprojecting the thermal Sunyaev-Zeldovich effect. Finally, we compare our results to a tree-level perturbation theory calculation and find reasonable agreement between the model and measurement.

79 ASTRONOMY AND ASTROPHYSICS↗

Evaluation of GlassNet for physics-informed machine learning of glass stability and glass-forming ability

Glassy materials form the basis of many modern applications, including nuclear waste immobilization, touch-screen displays, and optical fibers, and also hold great potential for future medical and environmental applications. However, their structural complexity and large composition space make design and optimization challenging for certain applications. Of particular importance for glass processing and design is an estimate of a given composition's glass-forming ability (GFA). However, there remain many open questions regarding the underlying physical mechanisms of glass formation, especially in oxide glasses. It is apparent that a proxy for GFA would be highly useful in glass processing and design, but identifying such a surrogate property has proven itself to be difficult. While glass stability (GS) parameters have historically been used as a GFA surrogate, recent research has demonstrated that most of these parameters are not accurate predictors of the GFA of oxide glasses. Here, in this work, we explore the application of an open-source pre-trained neural network model, GlassNet, that can predict the characteristic temperatures necessary to compute GS with reasonable performance and assess the feasibility of using these physics-informed machine learning (PIML)-predicted GS parameters to estimate GFA. In doing so, we track the uncertainties at each step of the computation—from the original ML prediction errors to the compounding of errors during GS estimation, and finally to the final estimation of GFA. While GlassNet exhibits reasonable accuracy on all individual properties, we observe a large compounding of error in the combination of these individual predictions for the PIML prediction of GS, finding that random forest models offer similar accuracy to GlassNet. We also break down the performance of GlassNet on different glass families and find that the error in GS prediction is correlated with the error in crystallization peak temperature prediction. Lastly, we utilize this finding to assess the relationship between top-performing GS parameters and GFA for two ternary glass systems: sodium borosilicate and sodium iron phosphate glasses. We conclude that to obtain true ML predictive capability of GFA, significantly more data needs to be collected.

36 MATERIALS SCIENCE↗

Calibration of RAFM Micromechanical Model for Creep Using Bayesian Optimization for Functional Output

A Bayesian optimization procedure is presented for calibrating a multimechanism micromechanical model for creep to experimental data of F82H steel. Reduced activation ferritic martensitic (RAFM) steels based on Fe(8–9)%Cr are the most promising candidates for some fusion reactor structures. Although there are indications that RAFM steel could be viable for fusion applications at temperatures up to 600°C, the maximum operating temperature will be determined by the creep properties of the structural material and the breeder material compatibility with the structural material. Due to the relative paucity of available creep data on F82H steel compared to other alloys such as Grade 91 steel, micromechanical models are sought for simulating creep based on relevant deformation mechanisms. As a point of departure, this work recalibrates a model form that was previously proposed for Grade 91 steel to match creep curves for F82H steel. Due to the large number of parameters (9) and cost of the nonlinear simulations, an automated approach for tuning the parameters is pursued using a recently developed Bayesian optimization for functional output (BOFO) framework (Huang et al., 2021, “Bayesian optimization of functional output in inverse problems,” Optim. Eng., 22, pp. 2553–2574). Incorporating extensions such as batch sequencing and weighted experimental load cases into BOFO, a reasonably small error between experimental and simulated creep curves at two load levels is achieved in a reasonable number of iterations. In conclusion, validation with an additional creep curve provides confidence in the fitted parameters obtained from the automated calibration procedure to describe the creep behavior of F82H steel.

42 ENGINEERING↗

Emerging Flexible Designs for Geospatial Multimodal Foundation Models

Foundation models are rapidly transforming Earth observation by enabling scalable pretraining across diverse unlabeled geospatial modalities. However, their architectural diversity—ranging from encoder-only to encoder-decoder and masked autoencoding paradigms—makes it challenging to assess performance trade-offs in a consistent manner. In this work, we present an apples-to-apples comparison of leading FM architectures designed for geospatial multimodal reasoning, with a particular focus on flexibility across varied spectral band configurations. We standardize pretraining using identical self-supervised learning objectives and training datasets, and evaluate all models under consistent parameterization on the GEOBench benchmark across classification and segmentation tasks. Our results offer new insights into the design trade-offs between model flexibility, modality alignment, and downstream task performance. By highlighting architectural strengths and limitations under controlled conditions, this study provides practical guidance for building next-generation geospatial foundation models capable of robust multimodal reasoning.

Ambrozio Dias, Philipe [ORNL] (ORCID:0000000194277↗

Electrolyte and Cutoff Potential Effects on Cycle Life of Li4Ti5O12/LiNi0.9Mn0.1O2 Batteries for Behind-the-Meter Storage Applications

Behind-the-Meter Storage (BTMS) is a stationary battery energy storage system that is connected to the electrical distribution system on the customer's side of the utility's service meter. BTMS systems are used to store electrical energy from the grid as well as inconstant, renewable energy, such as local solar and wind generation. A successful BTMS system will allow the customer to pair their energy generation and storage to optimize electrical consumption from the grid, improving reliability and minimizing cost. For BTMS applications, batteries must be designed and optimized with different set of criteria from other leading segments of the Li-ion battery market, like transportation, due the system being stationary and proximal to the residential or commercial building it's benefitting. BTMS applications prioritize safety, cost (low/no-critical materials), reliability (20-year calendar life), and durability (10,000 cycle life), while having the ability to (minimally) compromise energy density and rate capability. Lithium titanate (Li4Ti5O12-, LTO) is a promising anode candidate for BTMS applications due to its high safety and capacity retention, while maintaining a reasonable 160 mAhg-1 reversable capacity and composition of relatively abundant materials. (1) Specifically, LTO has a high working voltage which helps to prevent Li dendrite formation, improving safety. Furthermore, LTO also has negligible lithiation-based volume change, leading to less mechanical pulverization, or loss of active material, upon cycling. For the cathode, materials with little or no Co are of high interest due to the high cost and low abundance of Co. LiMn2O4 (LMO) has been paired with LTO for BTMS applications in the past due to its safety, low cost (abundancy), and reasonably high operating voltage. (2-4) However, the low capacity of LMO limits energy density and specific energy. While not the highest priority for BTMS applications, increasing energy density will enable deployment in space constrained BTMS applications and decrease total cost. LiNi0.9Mn0.1O2 (LN-MO) is a recently developed material with promise due to its high operating voltage and relatively low price. (5) However, Ni-rich layered oxides, including LNMO, tend to struggle with capacity retention during high-voltage cycling due to mechanical pulverization, irreversible phase transitions, and unstable solid-electrolyte interphase. The study presented here focuses on building an understanding of how electrolyte solvent and varied cutoff potentials will impact the cycle life of LTO/LN-MO cells. Specifically, a comparison is provided between ethylene carbonate (EC), ethyl methyl carbonate (EMC), fluoroethylene carbonate (FEC), and Gen2 electrolyte solvents with 1M Lithium hexafluorophosphate (LiPF6) salt, cycling to two upper termination potentials, 2.6V and 2.7V. Electrochemical testing and diagnostics (e.g., differential capacity analysis, area specific impedance, constant voltage hold, and rate capability) and post-mortem characterization will be used to understand the aging behavior and failure mechanisms of the 8 cell combinations (four electrolytes and two voltage cutoffs). Cells with FEC electrolyte showed a lower initial capacity compared to cells with Gen2, EMC, and EC cycling at both voltages; however, the cells with FEC showed consistent trends in capacity retention with 2.6V and 2.7V termination potentials, while the cells with the other electrolytes showed much higher rates of capacity loss when cycling to the higher voltage. These results indicate that FEC may play a role in improving durability of high-voltage, Ni-rich electrode systems for use in high-cycle applications, such as BTMS.

electrolyte↗

AFUE Analysis Software Tool

The Annual Fuel Utilization Efficiency (AFUE) analysis of residential and light commercial furnaces follows ANSI/ASHRAE Standard 103-2017 (i.e., Method of Testing for Annual Fuel Utilization Efficiency of Residential Central Furnaces and Boilers) . The analysis is a complex, comprehensive method based on furnace configuration and specific components equipped, requiring detailed furnace testing and measurement data. For this reason, an AFUE Analysis Tool using Microsoft Excel enabled with Visual Basic for Applications (VBA) was developed with a user-friendly interface and comprehensive coverage. The tool consists of three worksheets: unit and configuration selection, geometry and measurement data input, and AFUE plus key results. This tool can be used to estimate the AFUE of both condensing and noncondensing furnaces with single-stage, two-stage, and step-modulating functions. The tool was validated with experimental data from Oak Ridge National Laboratory’s natural gas furnace projects that are commercially available. The results indicate the tool is reasonably accurate in the evaluation of a new R&D modified furnace unit.

Gao, Zhiming [Oak Ridge National Laboratory (ORNL)↗

Identifying Climate Patterns Using Clustering Autoencoder Techniques

Abstract The complexity of growing spatiotemporal resolution of climate simulations produces a variety of climate patterns under different projection scenarios. This paper proposes a new data-driven climate classification workflow via an unsupervised deep learning technique that can dimensionally reduce the vast volume of spatiotemporal numerical climate projection data into a compact representation. We aim to identify distinct zones that capture multiple climate variables as well as their future changes under different climate change scenarios. Our approach leverages convolutional autoencoders combined with k -means clustering (standard autoencoder) and online clustering based on the Sinkhorn–Knopp algorithm (clustering autoencoder) across the conterminous United States (CONUS) to capture unique climate patterns in a data-driven fashion from the Geophysical Fluid Dynamics Laboratory Earth System Model with GOLD component (GFDL-ESM2G). The developed approach compresses 70 years of GFDL-ESM2G simulation at 0.125° spatial resolution across the CONUS under multiple warming scenarios to a lower-dimensional space by a factor of 660 000 and then tested on 150 years of GFDL-ESM2G simulation data. The results show that five climate clusters capture physically reasonable and spatially stable climatological patterns matched to known climate classes defined by human experts. Results also show that using a clustering autoencoder can reduce the computational time for clustering by up to 9.2 times when compared to using a standard autoencoder. Our five unique climate patterns resulting from the deep learning–based clustering of the lower-dimensional space thereby enable us to provide insights on hydrometeorology and its spatial heterogeneity across the conterminous United States immediately without downloading large climate datasets. Significance Statement This paper presents a data-driven climate classification approach using unsupervised deep learning to dimensionally reduce climate model outputs and to identify distinct climate regions for their future changes. Our approach compresses climate information for 70 years of Geophysical Fluid Dynamics Laboratory Earth System Model data across the conterminous United States (CONUS) at 0.125° spatial resolution. The results reveal that five climate clusters capture reasonable and stable climatological patterns matched to known climate patterns. The embedded clustering process in deep learning provides ×9.2 times faster execution than the k -means clustering technique. These results give us insight about climate spatial patterns and heterogeneity of hydrological patterns across the conterminous United States without downloading large climate datasets.

Kurihana, Takuya↗

An Observational Evaluation of RKW Theory over the U.S. Southern Great Plains

The theory of Rotunno et al. (“RKW” theory) addresses the behavior of squall-line cold pools in vertically sheared flows. It predicts that, within a given thermodynamic environment, a balance between baroclinic vorticity generation by the cold pool and low-level environmental vertical wind shear induces an upright updraft along the gust front that maximizes the initiation of new convective cells. Although this theory has been evaluated numerically, its applicability to observed systems remains unclear and is limited by a lack of critical measurements, including high-frequency thermodynamic and wind profiles across the gust front. Herein, observations from the Atmospheric Radiation Measurement Southern Great Plains (ARM-SGP) observatory near Lamont, Oklahoma, are used to evaluate RKW theory for 10 well-observed squall lines over a 11-yr period. For this evaluation, RKW parameters including cold-pool intensity (c), low-level ambient, line-normal vertical shear (ΔV n ), subcloud and cloud-layer updraft tilts, and multiple measures of system intensity are estimated. Furthermore, the c estimates rely on thermodynamic retrievals from the Atmosphere Emitted Radiance Interferometer (AERI), which are uncertain but verify reasonably well against independent observations. As predicted by the theory, for c/ΔV n ≥ 1, c/ΔV n correlates positively with updraft tilt and negatively with system intensity, but these results are not always statistically significant and are also sensitive to the method by which ΔV n is evaluated. Specifically, ΔV n evaluations that extend above the cold-pool top yield greater consistency with RKW predictions. Also, some measures of intensity correlate more strongly with standard moist instability metrics than with RKW parameters.

Cold pools↗

Understanding the Biases in Global Monsoon Simulations from the Perspective of Atmospheric Energy Transport

Understanding global monsoon (GM) variability and projecting its future changes rely heavily on climate models. However, climate models generally show pronounced biases in GM simulations, and the reasons for this remain unclear. Here, in this study, we evaluate the performance of 20 pairs of climate models that participated in both phase 5 of the Coupled Model Intercomparison Project (CMIP5) and phase 6 of CMIP (CMIP6) and identify the sources of their GM simulation biases from an energy transport perspective. The multimodel mean improvement in CMIP6 compared to CMIP5 is demonstrated by the increasing skill scores for various GM metrics from 0.20–0.79 to 0.48–0.83. More specifically, the dry biases in the Northern Hemisphere Summer Monsoon (NHSM) precipitation in CMIP5 [root-mean-square error (RMSE): 1.85 mm day −1 ] are reduced in CMIP6 (RMSE: 1.66 mm day −1 ). This higher simulation skill is associated with higher skill in simulating the precipitation-solstitial mode, monsoon intensity, and monsoon domains. The improvement in the NHSM precipitation simulation results from that in the meridional transport of atmospheric energy. Atmospheric energy budget analysis shows that the negative biases in downward surface longwave radiation and northward energy transport are smaller in CMIP6 than in CMIP5 in the boreal summer, resulting in a more realistic interhemispheric thermal contrast and meridional gradient of moist static energy. However, a major weakness of the CMIP6 models is found in the Southern Hemisphere Summer Monsoon precipitation simulation due to the positive bias in the top-of-the-atmosphere downward longwave radiation. This study shows that reasonably reproducing the meridional global atmospheric energy transportation is necessary for skillful GM simulation.

54 ENVIRONMENTAL SCIENCES↗

Exploring Flood Predictability in Taiwan through Coupled Atmospheric–Hydrological and High-Performance Hydrodynamic Models

Effective flood simulation capabilities can tremendously support early warning and disaster prevention. To examine the applicability of a fully physics-based and high-performance flood simulation and forecasting modeling framework for a flood-prone region in Taiwan, we conduct a numerical experiment that couples the Weather Research and Forecasting (WRF) Model, WRF-Hydrological modeling system (WRF-Hydro), and the Two-Dimensional Runoff Inundation Toolkit for Operational Needs (TRITON) to perform integrated rainfall, streamflow, and flood simulations. Furthermore, we first use the coupled WRF and WRF-Hydro (WWH) to predict rainfall and streamflow and then drive TRITON with the predicted streamflow hydrographs to simulate flood depth and inundation area. With the refined spatial resolution and parameterization, this framework can better predict rainfall with reasonable spatial patterns. Although WWH could overestimate the amount of rainfall in some areas, the uncertain rainfall–streamflow predictions produce reasonable flood maps able to pinpoint regions at risk of flooding. In terms of model efficiency, the graphics processing unit–based computation can yield a speed-up factor as high as ∼13 compared to the central processing unit–based computation, promoting the efficacy of the coupled modeling framework in practical real-time flood forecasting.

Coupled models↗

Data for Greenhouse Gas Accounting Procedures in Low Carbon Fuel Policies Overlook the Spatial Variability of Miscanthus-Derived Sustainable Aviation Fuel

Low carbon fuel policies such as the U.S. Renewable Fuel Standard (RFS), Canada Clean Fuel Regulations (CFR), and California Low Carbon Fuel Standard (LCFS) as well as the 45Z tax credit are intended to reduce greenhouse gas (GHG) emissions from transportation. Cellulosic feedstocks, optimized biorefineries, and favorable farming locations can significantly reduce biofuel carbon intensity (CI). Despite advances in field-to-fuel GHG monitoring and flexibility in resource allocation within biorefineries (e.g., governing net electricity production), rigid CI accounting procedures in current policies may limit CI responsiveness across candidate sites and processing facilities. This work examines a hypothetical biomass-to-sustainable aviation fuel (SAF) pathway using miscanthus and alcohol-to-jet (i) to demonstrate how GHG accounting requirements drive estimates of biofuel CIs and (ii) to explore potential CI and financial implications of scenario-specific life cycle assessment (LCA). Results demonstrate that GHG accounting using the CFR/LCFS can reasonably account for distinct levels of net electricity production by a biorefinery, but only the CFR yields similar CI sensitivity to spatially explicit factors (feedstock CI, grid electricity CI) as scenario-specific LCA: most GHG accounting frameworks do not capture CI variation across candidate sites in the United States. Ultimately, this work demonstrates the importance of LCA methodological specifications in low carbon fuel policies and tax credits.

Miscanthus↗

Data for Engineering and Evolution of Yarrowia lipolytica for Producing Lipids from Lignocellulosic Hydrolysates

Yarrowia lipolytica , an oleaginous yeast, shows promise for industrial fermentation due to its robust acetyl-CoA flux and well-developed genetic engineering tools. However, its lack of an active xylose metabolism restricts the conversion of cellulosic sugars to valuable products. To address this, metabolic engineering, and adaptive laboratory evolution (ALE) were applied to the Y. lipolytica PO1f strain, resulting in an efficient xylose-assimilating strain (XEV). Whole-genome sequencing (WGS) of the XEV followed by reverse engineering revealed that the amplification of the heterologous oxidoreductase pathway and a mutation in the GTPase-activating protein gene (YALI0B12100g) might be the primary reasons for improved xylose assimilation in the XEV strain. When a sorghum hydrolysate was used, the XEV strain showed superior xylose consumption and lipid production compared to its parental strain (X123). This study advances our understanding of xylose metabolism in Y. lipolytica and proposes effective metabolic engineering strategies for optimizing lignocellulosic hydrolysates.

Hydrolysate↗

Two-stage formation-energy correction (NbZr, TaZr, VZr)

This bundle contains the scripts, the raw and corrected per-structure data, and the manuscript plots for the NbZr / TaZr / VZr BCC binary formation energies and the associated RMSDs. Why a two-stage correction is necessary: The "raw" formation energy of every relaxed VASP configuration is computed in the usual way, FE_raw(c) = E_alloy(c) - sum_i x_i * E_pure_i , where E_pure_i are the per-atom total energies of the pure-element reference structures (Nb, Ta, V, Zr in the same BCC supercell, with identical INCAR / KPOINTS / PAW choices). With perfectly consistent reference runs the raw FE should vanish at the two pure-element endpoints (x = 0 and x = 1) by construction. In practice this does not hold for two reasons that are present in our dataset: 1. Reference-energy inconsistency (composition-dependent bias). Even with identical input parameters, the pure-element runs (stored in `corrected_DFT_pure_element_runs/`) differ slightly from the values that would be implied by the alloy runs at near-pure compositions (a few meV/atom). This bias is approximately linear in concentration, because the residual error in E_pure_Nb (or E_pure_Ta / E_pure_V) propagates into FE_raw(c) as (1 - x) * dE_pure_1, and the corresponding error in E_pure_Zr propagates as x * dE_pure_2. Left uncorrected, this produces a non-physical "tilt" of FE_raw(x) and shifts the entire FE-vs-x cloud away from zero at the endpoints. 2. Endpoint anchoring against the audited true endpoints. The strict endpoint values (FE_x0_meVatom, FE_x1_meVatom in `corrected_fe_strict_endpoints_20260518/strict_endpoint_check_20260518.csv`) were re-derived from an independent cross-check of the pure-element runs. After stage 1 removes the linear bias, the near-pure compositions in the alloy dataset still extrapolate to values that differ slightly from these audited endpoints — because stage 1 is fit from a few near-end alloy bins, not from the audited pure-element references themselves. The README.txt file discusses how these issues are addressed by the two-stage correction, and describes folder layout, pipeline summary, and how to re-run.

36 MATERIALS SCIENCE↗

An improved dataset for predicting mammal infecting viruses from genetic sequence information

There have been several attempts to develop machine learning (ML) models to identify human infecting viruses from their genomic sequences, with varying degrees of success. Direct comparison between models is problematic, because these models are typically trained and evaluated on different datasets with alternative data splitting schemes, features, and model performance metrics. In this paper we present a standardized dataset of mammal infecting and non-infecting viral pathogens, refined from the previous work of Mollentze et al. to include the latest literature evidence, roughly doubling the number of curated host-virus records available to the community, and new host target labels, primate and mammal. The new host labels were included for several reasons, including previous reports that classification performance is better at broader taxonomic ranks and the idea that there may be more data for primate infection that might serve as a suitable proxy for zoonotic potential and avoidance of false positives for human infection due to absence of evidence. On this dataset, we report the performance of eight machine learning models for predicting mammal-infecting viruses from their genomic sequences. We find that randomly assigning cases in our improved dataset to training/testing sets, when compared to the original assignments into training/testing in Mollentze et al., increases the overall average ROC AUC of prediction of human infection from 0.663 ± 0.070 to 0.784 ± 0.013, consistent with the reduction in phylogenetic distance between train and test sets (relative entropy change from 3.00 to 0.08). The broadest host category of mammal infection can be predicted most reliably at 0.850 ± 0.020. We share our improved dataset and code to enable standardized comparisons of machine learning methods to predict human host infections. Overall, we have presented preliminary evidence that classification of virus host infection is more tractable at higher taxonomic ranks, that unsurprisingly reducing the phylogenetic distance between training and test sets can improve predictive performance, that peptide kmer features appear to be harmful to out of sample model performance, and we are left with the question of whether models for virus host prediction can reasonably be expected to perform well in out of sample scenarios given the likelihood that viruses do not share a common ancestor. Consistent with this concern, when the data is resampled such that there is no overlap between viral families in training and test sets (relative entropy > 24), models perform no better than random chance at prediction of human infection regardless of whether kmers are included (ROC AUC 0.50 ± 0.08) or not (ROC AUC 0.50 ± 0.04).

59 BASIC BIOLOGICAL SCIENCES↗

A plan to revitalize the domestic superconducting radio-frequency industry

Superconducting radio-frequency (SRF) cavities are essential building blocks of modern particle accelerators for scientific research, and they offer unique capabilities that could be transformative for commercial applications. Growth of the domestic SRF industry in North America has faced several challenges over the past decades, as most of the international demand for cavities was supplied by European vendors. This contribution provides a brief review of the domestic industrial vendor space, an outlook of the global demand for SRF cavities and an outline of the challenges leading to this supply chain deficiency. One of the main challenges towards establishing a robust domestic SRF industry has been the large uncertainty in the demand. Meanwhile, research and development activities to raise technical readiness of SRF accelerators for industrial use have continued and several potential markets are emerging that may offer a consistent and growing demand for SRF cavities. Finally, reasons and means of establishing and sustaining competitive domestic suppliers are described.

Accelerator Physics↗

Improved Creep Testing Approach for Bentonite EBS

In the United States, the nuclear reactor contributes 18.7 percent of the nation’s total electricity generation and produces a significant amount of spent nuclear fuel (SNF) and high-level waste. Because SNF is being stored for longer periods than initially envisioned, the U.S. Department of Energy Office of Nuclear Energy, Office of Spent Fuel and Waste Science and Technology is assessing the technical performance of the SNF storage systems after extended durations. The concept of a sustainable geologic repository for SNF disposal relies heavily on the safety provided by multiple barriers, such as repository host rock, overlying rock formation, and human-made engineered barrier systems (EBS). One of the major components of EBS is bentonite-based buffer materials that isolate the nuclear-waste canisters from the host rock and fill the void left in the horizontally drilled boreholes. The resaturated bentonite has a swelling characteristic, which is useful for self-sealing the microcracks in the host rock and supporting the canister’s heavy weight. However, factors such as nonuniformity in the saturation, swelling pressure of bentonite, and the high temperature at the canister-bentonite interface may lead to instability of the waste canister inside the borehole. Therefore, evaluating the long-term deformation of bentonite-based material is necessary. RESPEC Company, LLC recently completed a preliminary study for the U.S. Department of Energy (DE SC0022804) that is attempting to understand the time-dependent deformation in the consolidated sand/bentonite (SB) specimens through triaxial creep experiments and improve the understanding of EBS performance under anticipated repository conditions. The project included developing a standard testing procedure for preparing the consolidated core specimen from a mixture of sand and bentonite in a 50:50 ratio by weight, fabricating a new pressure vessel in-house, modifying existing creep equipment to perform triaxial creep experiments on consolidated SB specimens, and performing six long-term triaxial creep experiments under low-deviatoric stress at two different temperature conditions, which is representative of hypothetical conditions that might be encountered in a geologic repository within a reasonable rate of success in completing creep experiments. The consolidated SB specimen was prepared using the novel specimen preparation procedure and had physical characteristics (e.g., moisture content and bulk density) reasonably comparable with the properties of bentonite-based buffer, which is proposed to be used in the EBS in the Swedish KBS-3H design and the Swiss design of SNF disposal in the horizontal drifts. The newly fabricated pressure vessel could withstand the confining pressure of up to 15 megapascals (MPa) while maintaining the test temperature of up to 90 degrees Celsius (°C) for long-term creep experiments. The preliminary results from the triaxial creep experiments suggested that the consolidated SB specimen may experience time-dependent deformation at a low-deviatoric stress state and room temperature condition. Based on the deformation recorded in the SB specimen from axial and radial linear variable differential transformers (LVDTs) and physical characteristic changes, it is apparent that the time-dependent deformation is a combination of consolidation and creep, which is influenced by the level of deviatoric stresses and temperature. The systematic creep testing of bentonite-based specimens helped develop a standard testing procedure applicable for analyzing the similar characteristics of shale or clay-like materials in different geologic conditions outside repository science, including in civil engineering and infrastructure projects, underground and open pit mining, and deep drilling.

58 GEOSCIENCES↗