Search NASA⌕ Search

SEARCH · Search NASA

Results for “predictive modeling”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 415 records · Page 23

Direct Air Capture Using Aqueous Amino Acid Solvents in a Crossflow Absorber

Carbon dioxide (CO 2 ) is the most abundant of all greenhouse gases (GHGs). CO 2 levels in the atmosphere are 50% higher than in the preindustrial era, trapping heat. CO 2 removal from the atmosphere by direct air capture (DAC) is needed to achieve the internationally agreed global temperature goals. The most common CO 2 capture technology is absorption by aminebased solvents in packed columns. Amino acid solutions have recently gained attention due to their advantages over traditional amine solvents. To be implemented effectively, DAC industrial processes need to handle large airflow rates in separation absorbers. The large air and solvent flow rates preclude the use of countercurrent columns due to high-pressure drops and the occurrence of flooding. Crossflow air−liquid absorbers are used to handle large air and liquid volumes due to their lower-pressure drop. The objective of this work is to study the influence of crossflow absorber geometric parameters and operating conditions on product formation and process efficiency. An already derived theoretical model for countercurrent absorbers has been modified to simulate the operation of a crossflow DAC absorber. The predictive model was implemented into a computer code that was used to study the efficiency of the processes as geometric equipment dimensions and operating parameters vary. Practical suggestions are made to design more efficient DAC processes.

Absorption↗

Large spatiotemporal variability in aerosol properties over central Argentina during the CACTI field campaign

Abstract. Few field campaigns with extensive aerosol measurements have been conducted over continental areas in the Southern Hemisphere. To address this data gap and better understand the interactions of convective clouds and the surrounding environment, extensive in situ and remote sensing measurements were collected during the Cloud, Aerosol, and Complex Terrain Interactions (CACTI) field campaign conducted between October 2018 and April 2019 over the Sierras de Córdoba range of central Argentina. This study describes measurements of aerosol number, size, composition, mixing state, and cloud condensation nuclei (CCN) collected on the ground and from a research aircraft during 7 weeks of the campaign. Large spatial and multiday variations in aerosol number, size, composition, and CCN were observed due to transport from upwind sources controlled by mesoscale to synoptic-scale meteorological conditions. Large vertical wind shears, back trajectories, single-particle measurements, and chemical transport model predictions indicate that different types of emissions and source regions, including biogenic emissions and biomass burning from the Amazon and anthropogenic emissions from Chile and eastern Argentina, contribute to aerosols observed during CACTI. Repeated aircraft measurements near the boundary layer top reveal strong spatial and temporal variations in CCN and demonstrate that understanding the complex co-variability of aerosol properties and clouds is critical to quantify the impact of aerosol–cloud interactions. In addition to quantifying aerosol properties in this data-sparse region, these measurements will be valuable to evaluate predictions over the midlatitudes of South America and improve parameterized aerosol processes in local, regional, and global models.

54 ENVIRONMENTAL SCIENCES↗

Cloud Feedback Uncertainty in the Equatorial Pacific Across CMIP6 Models

Cloud feedback is the largest uncertainty in estimating Equilibrium Climate Sensitivity. In this study we focus on the equatorial Pacific, where CMIP6 model cloud feedback spread is notably large. Cloud radiative effects in this region are relevant for the global climate. Our findings show that models predict a consistent shift towards the ascent regime in response to El Nino-like sea surface warming. Models diverge in terms of the radiative impact due to differences in cloud characteristics in ascent and subsidence regimes. Using the observed relationship between circulation regime and cloud radiative effect, we find a reduction in the regional mean cloud feedback estimate from 0.77 to 0.22 W m -2 K -1 , though this does not substantially lessen the model spread in total feedback. Pathways to reduce this spread include: improving confidence in the regional ocean warming pattern, and using observations and models to understand cloud type and circulation interactions.

CMIP6↗

EASY-SHIFT v Alpha

The software is a generic, price- and load-responsive control algorithm integrating heat pumps with thermal energy storage. The algorithm leverages simple models of the system and easily accessible data to schedule operation of heat pumps and thermal energy storage in ways that minimize the cost of operating the heating/cooling system. This tool is specifically designed to be easy to interact with, and something that industry partners are able to adopt. There are two current state of the art approaches. Industry tends to develop very simple algorithms, with predetermined schedules that are not capable of changing operation in response to changes in operating environment. For example, a control designed to avoid high-price electricity from 5-8 PM will not be able to adapt if the high-price period changes to 4-9 PM. Academia commonly develops algorithms called Model predictive control (MPC). MPC requires extensive data and highly trained staff to develop a specific type of simulation model of the building, connect the building to optimization algorithms, and leverage powerful computers. Industry, with limited time/finance budgets for any project, is resistant to adopting MPC due to the associated high complexity and cost.

Grant, Peter [Lawrence Berkeley National Laborator↗

From soil to sequence: filling the critical gap in genome-resolved metagenomics is essential to the future of soil microbial ecology

Abstract Soil microbiomes are heterogeneous, complex microbial communities. Metagenomic analysis is generating vast amounts of data, creating immense challenges in sequence assembly and analysis. Although advances in technology have resulted in the ability to easily collect large amounts of sequence data, soil samples containing thousands of unique taxa are often poorly characterized. These challenges reduce the usefulness of genome-resolved metagenomic (GRM) analysis seen in other fields of microbiology, such as the creation of high quality metagenomic assembled genomes and the adoption of genome scale modeling approaches. The absence of these resources restricts the scale of future research, limiting hypothesis generation and the predictive modeling of microbial communities. Creating publicly available databases of soil MAGs, similar to databases produced for other microbiomes, has the potential to transform scientific insights about soil microbiomes without requiring the computational resources and domain expertise for assembly and binning.

59 BASIC BIOLOGICAL SCIENCES↗

Development and Validation of Home Comfort System for Total Performance Deficiency/Fault Detection and Optimal Comfort Control

In this project, we developed and tested a learning-based home thermal model that facilitates the operation of a model predictive control (MPC)-based optimization agent and an automated fault detection and diagnosis (AFDD) agent. The home thermal model was constructed using a two-node resistor-capacitor model. Moreover, two accompanying parameter identification methods were introduced, least-squares and optimization. Based on the home thermal model, the MPC-based optimization agent was developed to optimize residential HVAC operation. Using two FDD methods, the AFDD agent was constructed to detect and diagnose two prevalent residential AC faults, airflow reduction and refrigerant undercharge. The home thermal model, along with the MPC-based optimization agent and AFDD agent, were tested at the Norman Test House, Miami Test House, Pacific Northwest National Laboratory (PNNL) Test House A, and PNNL Test House B. Finally, they were also field tested in nine demonstration homes with real occupants.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

A Generalized Grain-Scale Model for the Non-Plasma and Plasma-Assisted Hydrogen Direct Reduction of Iron Ore

Direct Reduction of Iron ore using hydrogen (H-DRI) is a promising pathway towards efficient steelmaking and accurate predictive models are a necessity for scale-up and optimization of this technology. However, accurate models of this process remain limited because existing models oversimplify grain-scale phenomena, such as nonlinearity inside grain, self-sufficient porosity, surface reactions, and the role of plasma species. These phenomena are important for flash steelmaking and plasma-assisted H-DRI processes. To address this need, we present a phenomenological model for simulating H-DRI at the scale of a single micron-sized grain of the iron ore. We call this the Transient Reactive Grain Model (TRGM). TRGM incorporates key physical process: gas species transport, a chemical kinetics of material conversion, nanopore structural evolution and, adsorption-desorption surface kinetics at the reactive nanopore surface. The important contribution of this work is that the model provides a dependence on different reductant species, specifically hydrogen atoms versus molecules, so that role of hydrogen plasma reduction can be clarified compared to the use of pure hydrogen gas reduction. TRGM predictions agree well with experimental data for both molecular H2 reduction of Fe2O3 and plasma hydrogen reduction of Fe3O4. Results reveal species concentration gradients with a diffuse reaction zone, and enhanced hydrogen diffusion at the grain outer surface due to evolving porosity. These findings challenge common assumptions in existing models, including sharp reaction fronts, quasi-steady diffusion and kinetics, and the neglect of surface chemistry. As a generalized grain-scale model for H-DRI processes, TRGM has practical applications in flash steelmaking and in-flight reduction using both molecular and plasma hydrogen.

08 HYDROGEN↗

Incorporating long-range dependence and fractal features in turbulence spectra

We introduce an advanced turbulence spectrum model developed from mathematical foundations from a covariance function class and empirically validated using extensive field data. This model captures the complex dynamics of long-range dependence, and fractal characteristics prevalent in riverine and atmospheric boundary layer (ABL) flows that are ignored by classical spectrum models, such as IEC (International Electrotechnical Commission) von Kármán and Kaimal model. The model delineates scaling behaviors across distinct frequency bands and offers substantial flexibility through five well-defined parameters each characterizing a distinct physical aspect of the velocity time series. A detailed procedure for obtaining each parameter from time series data is outlined. The comprehensive validations with field data from tidal currents and ABL flows substantiate the model’s fidelity in accurately replicating observed phenomena. This validation establishes the reliability of the proposed model and, when incorporated into stochastic full-field simulators such as TurbSim, demonstrates its potential to advance the predictive modeling and analysis of turbulent flows in environmental science and engineering contexts.

Cheng, Shyuan [Univ. of Illinois at Urbana-Champai↗

Diagnostics: Chapter 8 of the special issue: on the path to tokamak burning plasma operation

This chapter presents the activity conducted by the ITPA topical group (TG) on Diagnostics over about the last 15 years. Following a general introduction of the ITER Diagnostics led by their measurement roles, the document is organized in several subchapters detailing the design support, research and development activity conducted by each of the specialist working groups (WGs) of the TG. Please note that the magnetic diagnostics were supported at the TG without a specific WG. Their status is included in the general introduction. In the following some highlights of the subchapter’s contents are provided. Recent advances in ITER first wall (FW) diagnostics for the measurements of plasma-metallic wall interaction in support of the ITER research plan are reported. An InfraRed imaging Video Bolometer for ITER has been developed and tested on several tokamaks to measure the radiated power loss. A laser-induced breakdown spectroscopy (LIBS) technique which utilizes a pulsed laser beam to ablate locally by forming a crater, will measure local tritium inventory in the FW material. Real-time Residual Gas Analyzers will measure the neutral gas composition in a divertor port and an equatorial port during plasma operation. Due to the full metallic FW environment, the plasma-wall interaction in ITER will face several challenges such as the compromised radiated power and divertor heat flux measurements by reflection. Ray tracing and analysis codes have been developed to eliminate and correct the effects of reflection in the measurements. The characteristics of the reflecting surfaces depending on the roughness and angle of the incidence have been measured by dedicated experiments, and the results were applied to the reflection elimination. For the measurement of the metallic impurity radiation induced by eroded metallic atoms, a vacuum ultraviolet spectrometer has been developed and tested. An extensive thermonuclear diagnostic suite will be required to support the operation of ITER and the planned experimental program for future burning plasma experiments. Due to the harsh environmental conditions, the implementation of diagnostic systems in ITER is a major challenge. These conditions include high levels of neutron and gamma fluxes, neutron heating, particle bombardment. Therefore, the selection and design of diagnostic systems must take into account a number of phenomena previously unseen in diagnostic design. For this reason, the measurement of neutrons and confined or lost fast ions, with particular emphasis on alpha particles, is critical to ITER. The diagnostics associated with these measurements will be important for future plasma-burning experiments at ITER. The high neutron emission and very large plasma size in ITER make neutron diagnostics the main diagnostic method used to measure plasma parameters such as fusion power, fusion power density, ion temperature, energy of fast ions and their spatial distributions in the plasma core. Active spectroscopy techniques are methods where a neutral particle beam is injected into the plasma and information on plasma parameters is extracted from the measurement of line emission resulting from the beam-plasma interaction, either by plasma ions or by beam atoms. Spatial localization is achieved by crossing the beamline and multiple observation lines. The ITER plasma will be a high temperature, moderately dense, fully ionized collisional plasma. The plasma facing surfaces are principally metallic being fashioned from beryllium or tungsten but many other elements, arising from either structural or from operational needs, may enter this plasma. The energy range of the emitted photons range from meV (infra-red) to multi keV (x-rays) and originate from all areas of the plasma volume. The primary role of passive emission diagnostics is to identify what is in the plasma from spectral signatures. Extracting quantitative information from these measurements such as impurity content, ion temperature, rotation, degree of detachment and radiated power depends on calibrated instruments, a physics model of the atomic and molecular processes and plasma transport and an analysis workflow that takes into account environmental effects such as reflections. The particular needs for ITER have prompted a multi-machine, many-year effort to address all these aspects and this chapter reviews the work on diagnostic design, experiments and new analysis techniques. An overview of the laser diagnostics to be implemented on ITER is also provided in this paper. This includes descriptions of the Thomson scattering in the core, edge and divertor regions, polarimetry and interferometry diagnostics used for measuring plasma density and also measurements of helium density in the divertor using Laser Induced Flourescence. Techniques which can allow improvements on current measurements are also addressed in particular expanding poloidal polarimetry measurements to measure field fluctuations and proposed use of dispersion interferometery which has a number of advantages over existing methods. This paper identifies particular areas where further research and testing on existing tokamaks is useful even at this advanced stage to inform the design of diagnostics for ITER. Outstanding areas of concern for the implementation of laser diagnostics, in particular with a view to reliable operation are identified. An overview of the latest developments of microwave diagnostic systems and techniques is given. The primary focus is the contributions for ITER—the next step burning plasma experiment—which is supplemented by describing recent progress of techniques applicable for fusion experiments beyond ITER. The contributions are intentionally kept concise, and are being supplemented by a rich list of references for further studies. Radiation induced effects are receiving continuous and well-deserved attention of the ITER diagnostic community and they are in many cases one of the primary design drivers of the ITER diagnostic systems. The paper summarizes recent progress in this area focusing primarily on the ITER diagnostics but in some cases provides also outlook for the possible solutions for even more demanding radiation environment of fusion reactors beyond ITER. Despite advancements in the area of modeling and simulation of various radiation induced effects, experimental testing in a nuclear environment as close as possible to the target one is still seen as unavoidable for proper qualification of particular diagnostic functional elements. Recent advancement within three diagnostic areas: optical diagnostics, magnetics and bolometers is covered. Encouraging results on qualification of silica glass vacuum window assemblies are presented. In the area of magnetic sensors, progress of irradiation tests performed on ITER in-vessel LTCC inductive sensors is presented with outlook for novel technological approaches to inductive sensors utilizing thick printing and photolithography technologies being highlighted. Summary of advancements in the area of steady state magnetic field sensors based on Hall effect is given. New results of neutron irradiation test of the ITER borosilicate glass inserts for vacuum electrical feedthroughs are summarized finding negligible swelling at target level of neutron fluence. Off-line irradiation tests of fiber optic current sensors for plasma current measurement demonstrated that both for gamma doses up to 5 MGy and a total neutron fluence up to 10 15 cm −2 , radiation induced changes are still compatible with required measurement accuracy on ITER. The ITER bolometers are given as an example how considering radiation effects may influence the diagnostic design. Finally, outlook for future main R&D directions is outlined. All optical and laser-based diagnostics in ITER will be using mirrors to guide plasma radiation toward detectors, cameras and sensors. In the hostile plasma, radiation and particle environment the optical characteristics of diagnostic mirrors will degrade directly affecting the entire performance of involved diagnostic systems. An assessment of factors affecting mirror performance is provided. Among the prime adverse factors are deposition of plasma impurities, sputtering of mirror surface and steam ingress in the vicinity of mirrors. Within the International Tokamak Physics Activity with active support by ITER central team and domestic agencies, the structured research and development (R&D) program on mitigation of risks for diagnostic mirrors is underway. Within this program the mirror material development, the passive mitigation of mirror degradation by using diagnostic ducts and shutters along with an active mirror recovery program comprising the in-situ mirror cleaning and calibration is underway. Recent developments in diagnostic mirror R&D are described in this Chapter along with an example of their implementation of R&D solutions in ITER Infrared Thermography diagnostic. An assessment of still open engineering and physics questions, considerations on mirror risks during an early phase of ITER operation are given along with an overview of diagnostic mirror evolution in the late ITER operation stage toward the demonstration fusion power plant. Several crucial areas of diagnostic R&D outlined in ITER Research Plan are addressed. The basic control groups in a fusion reactor can be broken-down in five categories: (1) plasma position, magnetic configuration, and plasma current control, (2) profile control and confinement optimization, (3) MHD control and suppression, (4) edge dissipation control, radiation and plasma exhaust control and (5) break-down optimization. These categories are coupled via the physics (a control action in one domain will affect the other domains) and via shared actuators (e.g. ECRH for impurity accumulation avoidance, current density distribution control and MHD suppression). Consequently, a supervisory control system should determine the priority of the various control tasks, their couplings, and the interfaces with the safety and interlock system. For the systematic development of the various controllers taking the complexity of the plasma and the control system into account, a model-based approach is required. A short historical overview is given of the developments in systems and control theory and control engineering with special emphasis on those developments that are most relevant for Nuclear Fusion research and operation. An overview is given of the state of the field of fusion plasma control for the control categories. It will be shown how synthetic diagnostics are being developed in ITER and how they are used in diagnostic design and design validation and how they can be in model-based controller synthesis using relatively simple models. In modern control methods, multiple diagnostics are used to constrain relatively simple models. The constrained models provide an estimate for the state. This opens the route to state controllers, such as model predictive control. A major challenge in nuclear fusion research is the coherent combination of data from heterogeneous diagnostics and modeling codes for machine control and safety as well as physics studies. Measured data from different diagnostics often provide information about the same subset of physical parameters. Additionally, information provided by some diagnostics might be needed for the analysis of other diagnostics. A joint analysis of complementary and redundant data allows, e.g. to improve the reliability of parameter estimation, to increase the spatial and temporal resolution of profiles, to obtain synergistic effects, to consider diagnostics interdependencies and to find and resolve data inconsistencies. Physics-based modeling and parameter relationships provide additional information improving the treatment of ill-posed inversion problems. A coherent combination of all kind of available information within a probabilistic framework allows for improved data analysis results. The concept of integrated data analysis (IDA) in the framework of Bayesian probability theory is outlined and contrasted with conventional data analysis. Components of the probabilistic approach are summarized and specific ingredients beneficial for data analysis at fusion devices are discussed.

ITER↗

Measurement of muon neutrino induced charged current interactions without charged pions in the final state using a new T2K off-axis near detector WAGASCI-BabyMIND

We report a flux-integrated cross section measurement of muon neutrino interactions on water and hydrocarbon via charged current reactions without charged pions in the final state with the WAGASCI-BabyMIND detector, which was installed in the T2K near detector hall in 2018. The detector is located 1.5° off-axis and is exposed to a more energetic neutrino flux than ND280, another T2K near detector, which is located at a different off-axis position. The total flux-integrated cross section is measured to be 1.26 ± 0.18⁢(stat+syst) × 10 −39 cm 2 /nucleon on CH and 1.44 ± 0.21⁢(stat+syst) × 10 −39 cm 2 /nucleon on H 2 ⁡O. These results are compared to model predictions provided by the neut v5.3.2 and genie v2.8.0 Monte Carlo generators and the measurements are compatible with these models. Differential cross sections in muon momentum and cosine of the muon scattering angle are also reported. This is the first such measurement reported with the WAGASCI-BabyMIND detector and utilizes the 2020 and 2021 datasets.

lepton-hadron interactions↗

Half-Life and Precision Shape Measurement of the 2⁢𝜈⁢𝛽⁢𝛽 Decay of 130 Te

Here, we present a new measurement of the 2⁢𝜈⁢𝛽⁢𝛽 half-life of 130 Te (𝑇$^{2⁢𝜈}_{1/2}$) using the first complete model of the CUORE data, based on 1038 kg yr of collected exposure. Thanks to optimized data selection, we achieve a factor of two improvement in precision, obtaining 𝑇$^{2⁢𝜈}_{1/2}$ = (9.32⁢$^{+0.05}_{−0.04}$⁢stat⁢ $^{+0.07}_{−0.07}$⁢syst)×10 20 yr. The signal-to-background ratio is increased by 70% compared to our previous results, enabling the first application of the improved 2⁢𝜈⁢𝛽⁢𝛽 formalism to 130 Te . Within this framework, we determine a credibility interval for the effective axial coupling in the nuclear medium as a function of nuclear matrix elements. We also extract values for the higher-order nuclear matrix element ratios: second-to-first and third-to-first. The second-to-first ratio agrees with nuclear model predictions, while the third-to-first ratio deviates from theoretical expectations. These findings provide essential tests of nuclear models and key inputs for future 0⁢𝜈⁢𝛽⁢𝛽 searches.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Search for low-mass resonances decaying into two jets and produced in association with a photon or a jet at √s = 13 TeV with the ATLAS detector

A search is performed for localized excesses in the low-mass dijet invariant mass distribution, targeting a hypothetical new particle decaying into two jets and produced in association with either a high transverse momentum photon or a jet. The search uses the full Run 2 data sample from LHC proton-proton collisions collected by the ATLAS experiment at a center-of-mass energy of 13 TeV during 2015–2018. Two variants of the search are presented for each type of initial-state radiation: one that makes no jet flavor requirements and one that requires both of the jets to have been identified as containing b-hadrons. No excess is observed relative to the Standard Model prediction, and the data are used to set upper limits on the production cross section for a benchmark Z' model and, separately, for generic, beyond the Standard Model scenarios which might produce a Gaussian-shaped contribution to dijet invariant mass distributions. The results extend the current constraints on dijet resonances to the mass range between 200 and 650 GeV.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Data-driven prediction of scaling and ignition of inertial confinement fusion experiments

Recent advances in inertial confinement fusion (ICF) at the National Ignition Facility (NIF), including ignition and energy gain, are enabled by a close coupling between experiments and high-fidelity simulations. Neither simulations nor experiments can fully constrain the behavior of ICF implosions on their own, meaning pre- and postshot simulation studies must incorporate experimental data to be reliable. Linking past data with simulations to make predictions for upcoming designs and quantifying the uncertainty in those predictions has been an ongoing challenge in ICF research. We have developed a data-driven approach to prediction and uncertainty quantification that combines large ensembles of simulations with Bayesian inference and deep learning. The approach builds a predictive model for the statistical distribution of key performance parameters, which is jointly informed by past experiments and physics simulations. The prediction distribution captures the impact of experimental uncertainty, expert priors, design changes, and shot-to-shot variations. We have used this new capability to predict a 10× increase in ignition probability between Hybrid-E shots driven with 2.05 MJ compared to 1.9 MJ, and validated our predictions against subsequent experiments. We describe our new Bayesian postshot and prediction capabilities, discuss their application to NIF ignition and validate the results, and finally investigate the impact of data sparsity on our prediction results.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Knowledge Oriented Graph Unified Transformer (KOGUT) v0.1

KOGUT — Knowledge Oriented Graph Unified Transformer KOGUT implements the Relational Graph Transformer (RelGT) architecture for knowledge graph link prediction in biological domains, with a primary focus on microbial growth media prediction. While the original RelGT (arXiv:2505.10960) targets relational tables, time series, and multi-table databases, KOGUT adapts this architecture for heterogeneous biological knowledge graphs, providing first-in-class AI predictive models for microbial cultivation. Key Adaptations Beyond Original RelGT: - Knowledge Graph Focus: Applied to biological KGs with semantic node types (taxa, chemicals, media, phenotypes, environments) versus generic relational database tables, trained on the KG-Microbe knowledge graph (1.3M entities, 2.9M edges, 24 relation types). - Multimodal Node Encoding: Integrates node labels, categories, descriptions, and synonyms from KG metadata through learned embedding layers—adapting relational column features to graph node attributes with textual semantics. - Extended K-Hop Subgraph Strategy: Optimized neighborhood sampling (3-hop default, configurable up to 200 nodes) tuned for sparse biological networks, building on the original local-global attention framework with biological relation preservation. - Biolink Predicate Preservation: Type-specific transformations for 24 biological edge semantics (occurs_in, consumes, produces, has_phenotype, subclass_of) beyond standard relational foreign keys, enabling multi-relation link prediction. - Inductive Learning Support: Enables zero-shot predictions for novel taxa through feature-based embeddings (temperature, oxygen requirements, gram stain, cell shape), extending the original transductive relational benchmark scope to uncultured microorganisms. CheapSOTA Performance Optimizations (This Distribution): - VQ-EMA Centroid Attention: Vector quantization with exponential moving average for improved global context modeling (+5-10% MRR improvement). - HDF5 Precomputed Data Loading: One-time preprocessing of k-hop subgraphs to eliminate redundant graph traversals (2-5× training speedup). - Distributed Data Parallel Training: Multi-GPU support for scaling to larger knowledge graphs (tested on 4× NVIDIA A100 GPUs at NERSC Perlmutter). - Mixed Precision Training: Automatic mixed precision (AMP) for memory efficiency and faster training. Advantages Over Standard Knowledge Graph Embedding Models: Combines RelGT's proven multi-element tokenization (features, type, hop, structure) with graph-native biological representations, enabling interpretable link prediction across heterogeneous entities that standard embedding models (TransE, RotatE, ComplEx) and table-based transformers cannot directly model. Achieves near-perfect performance on microbial growth media prediction (MRR: 0.9966, Precision@1: 0.9932, Hit@10: 1.0000) while maintaining explainability through attention-based reasoning over biological pathways. Training Data: - KG-Microbe merged knowledge graph: 1,379,337 nodes, 2,960,472 edges - 24 biological relation types including taxonomic hierarchies, metabolic interactions, phenotype associations, and environmental relationships - Primary prediction task: Growth media suitability for microbial taxa (biolink:occurs_in, 50K edges) - Multi-relation capability: Predicts links for any of the 24 relation types, including chemical consumption/production, phenotype associations, and taxonomic classification Citation: Original RelGT Architecture: Dwivedi et al., "Relational Graph Transformer", arXiv:2505.10960, 2025 KOGUT Implementation: Knowledge Oriented Graph Unified Transformer for Microbial Growth Media Prediction Developed at Lawrence Berkeley National Laboratory (LBNL) Trained on NERSC Perlmutter supercomputer

Joachimiak, Marcin [Lawrence Berkeley National Lab↗

A Statistician’s Overview of Physics-Informed Neural Networks for Spatio-Temporal Data

The recent success of deep neural network models with physical constraints (so-called, Physics-Informed Neural Networks, PINNs) has led to renewed interest in the incorporation of mechanistic information in predictive models. Statisticians and others have long been interested in this problem, which has led to several practical and innovative solutions dating back decades. In this overview, we focus on the problem of data-driven prediction and inference of dynamic spatio-temporal processes that include mechanistic information, such as would be available from partial differential equations, with a strong focus on the quantification of uncertainty associated with data, process, and parameters. Here, we give a brief review of several paradigms and focus our attention on Bayesian implementations given they naturally accommodate uncertainty quantification. We then show that it is straight-forward to include the Bayesian PINN (B-PINN) within the Bayesian hierarchical model (BHM) framework that has long been considered for modeling dynamic spatio-temporal processes. Such a BHM-PINN is illustrated via a simulation study in which a latent nonlinear Burgers’ equation PDE governs the dynamics of Poisson distributed spatio-temporal data. Supplementary materials for this article are available online, including a standardized description of the materials available for reproducing the work.

Bayesian↗

Influence of particle size on NIR spectroscopic characterization of sorghum biomass for the biofuel industry

NIR spectroscopy is a rapid and accurate green technology for high-throughput biomass characterization, including sorghum (Sorghum bicolor), a promising energy crop for the biofuel industry. This study assessed the influence of particle size on NIR spectroscopic analysis (wavelength range: 867–2535 nm) of sorghum biomass composition. Grown under field conditions, a total of 113 types of genetically diverse sorghum accessions were dried, ground, and sieved (<250, 250–600, 600–850, and > 850 µm particle size) for developing partial least square regression (PLSR) prediction models for moisture, ash, extractive, glucan, xylan, acid-soluble lignin (ASL), acid-insoluble lignin (AIL), and total lignin (ASL + AIL). Overall, smaller particle sizes provided better model performance, while no single particle size provided the best performance for all the selected components. With only 9 selected bands and 4 latent variables (LVs), the best PLSR model was obtained for moisture with particle size of 600–850 µm with the square root of the coefficient of determination (R) of 0.85, the ratio of prediction to deviation (RPD) of 2.2, and the root mean square error (RMSE) of 0.46 % in external validation. Similar model performances were also obtained for ash, extractive, glucan, and xylan. This study showed that size reduction could effectively improve NIR spectroscopic analysis for lipid-producing sorghum biomass for the biofuel industry.

09 BIOMASS FUELS↗

Data for Influence of Particle Size on NIR Spectroscopic Characterization of Sorghum Biomass for the Biofuel Industry

NIR spectroscopy is a rapid and accurate green technology for high-throughput biomass characterization, including sorghum ( Sorghum bicolor ), a promising energy crop for the biofuel industry. This study assessed the influence of particle size on NIR spectroscopic analysis (wavelength range: 867–2535 nm) of sorghum biomass composition. Grown under field conditions, a total of 113 types of genetically diverse sorghum accessions were dried, ground, and sieved (<250, 250–600, 600–850, and > 850 µm particle size) for developing partial least square regression (PLSR) prediction models for moisture, ash, extractive, glucan, xylan, acid-soluble lignin (ASL), acid-insoluble lignin (AIL), and total lignin (ASL + AIL). Overall, smaller particle sizes provided better model performance, while no single particle size provided the best performance for all the selected components. With only 9 selected bands and 4 latent variables (LVs), the best PLSR model was obtained for moisture with particle size of 600–850 µm with the square root of the coefficient of determination (R) of 0.85, the ratio of prediction to deviation (RPD) of 2.2, and the root mean square error (RMSE) of 0.46 % in external validation. Similar model performances were also obtained for ash, extractive, glucan, and xylan. This study showed that size reduction could effectively improve NIR spectroscopic analysis for lipid-producing sorghum biomass for the biofuel industry.

Biomass Analytics↗

The Atacama Cosmology Telescope: DR6 constraints on extended cosmological models

We use new cosmic microwave background (CMB) primary temperature and polarization anisotropy measurements from the Atacama Cosmology Telescope (ACT) Data Release 6 (DR6) to test foundational assumptions of the standard cosmological model, ΛCDM, and set constraints on extensions to it. We derive constraints from the ACT DR6 power spectra alone, as well as in combination with legacy data from the Planck mission. To break geometric degeneracies, we include ACT and Planck CMB lensing data and baryon acoustic oscillation data from DESI Year-1. To test the dependence of our results on non-ACT data, we also explore combinations replacing Planck with WMAP and DESI with BOSS, and further add supernovae measurements from Pantheon+ for models that affect the late-time expansion history. We verify the near-scale-invariance (running of the spectral index dn s /d ln k = 0.0062 ± 0.0052) and adiabaticity of the primordial perturbations. Neutrino properties are consistent with Standard Model predictions: we find no evidence for new light, relativistic species that are free-streaming (N eff = 2.86 ± 0.13, which combined with astrophysical measurements of primordial helium and deuterium abundances becomes N eff = 2.89 ± 0.11), for non-zero neutrino masses (∑m ν < 0.089 eV at 95% CL), or for neutrino self-interactions. We also find no evidence for self-interacting dark radiation (N idr < 0.134), or for early-universe variation of fundamental constants, including the fine-structure constant (α EM /α EM,0 = 1.0043 ± 0.0017) and the electron mass (m e /m e,0 = 1.0063 ± 0.0056). Our data are consistent with standard big bang nucleosynthesis (we find Y p = 0.2312 ± 0.0092), the COBE/FIRAS-inferred CMB temperature (we find T CMB = 2.698 ± 0.016 K), a dark matter component that is collisionless and with only a small fraction allowed as axion-like particles, a cosmological constant (w = -0.986 ± 0.025), and the late-time growth rate predicted by general relativity (γ = 0.663 ± 0.052). We find no statistically significant preference for a departure from the baseline ΛCDM model. In fits to models invoking early dark energy, primordial magnetic fields, or an arbitrary modified recombination history, we find H 0 = 69.9 +0.8 -1.5 , 69.1 ± 0.5, or 69.6 ± 1.0 km/s/Mpc, respectively; using BOSS instead of DESI BAO data reduces the central values of these constraints by 1–1.5 km/s/Mpc while only slightly increasing the error bars. In general, models introduced to increase the Hubble constant or to decrease the amplitude of density fluctuations inferred from the primary CMB are not favored over ΛCDM by our data.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗