Search NASA⌕ Search

SEARCH · Search NASA

Results for “learning framework”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 361 records · Page 20

Coarse Graining Discrete Element Method Information in Particle-in-Cell Length Scales Using a Machine Learning Approach

This report details the development of a machine learning (ML)-driven framework to coarse-grain inter-particle collision dynamics from high-fidelity Discrete Element Method (DEM) simulations to Particle-in-Cell (PIC) scales for gas-solid systems. Traditional PIC models, while computationally efficient, rely on empirical granular stress formulations that fail to capture the full complexity of collision physics, particularly the heterogeneity in particle dynamics. This study adopts a bottom-up approach, integrating insights from DEM simulations to improve the physical fidelity and interpretability of PIC-scale models.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Integration of the Biot–Gassmann Fluid Substitution Method and Machine Learning-Based Velocity–Stress Relationship for Estimating In Situ Stresses

Recent advancements have shown that in situ stresses can be reliably estimated through an integrated machine/deep learning (ML/DL)-based framework, which relies on models trained and validated using true triaxial ultrasonic velocity (TUV) experimental data that involve measurements of ultrasonic velocity in saturated rocks under varying stress configurations. However, when the goal is to interpret lower frequency measurements, it may be more appropriate to run experiments on dry rocks and then obtain Biot–Gassmann-derived equivalent saturated velocities (low-frequency approximation) and employ these quantities for training ML/DL models to predict in situ stress. Whether the dispersion effect of frequency on the velocity–stress relationship substantially impacts in situ stress prediction is an important and unresolved question. This work presents an enhancement of ML/DL-based workflow by training and implementing ML/DL models using equivalent saturated acoustic velocities (low-frequency) obtained by applying Biot–Gassmann fluid substitution on the ultrasonic velocities of dry cores. The models were trained on TUV data sets derived from three subsurface cores extracted from the geothermal well 16B(78)-32 at the Utah FORGE site. Each core was subjected to 75 unique stress configurations for velocity measurement in the dry state. The ML/DL trained on the TUV data set with equivalent saturated velocities demonstrated promising performance to predict in situ stress in subsurface geological rocks using velocity–stress relationships with R 2 of 0.86, 0.971, and 0.975 and root mean squared error (RMSE) of 2.59, 1.92, and 1.80 for validation/testing phases of vertical, minimum horizontal, and maximum horizontal stress models, respectively. Additionally, interpretation and explanation by Shapley additive explanations (SHAP) analysis further improved scientific validation and model reliability for estimating in situ stresses.

colloids↗

Emerging anomaly detection techniques for electronic health records: A survey

Background Anomaly detection in electronic health records (EHRs) is a cornerstone of biomedical informatics, with direct implications for patient safety, clinical decision-making, and the prevention of healthcare fraud. Once guided primarily by simple rule-based methods, the field has advanced rapidly, driven by increased computing power, richer and more detailed health data, and the rise of machine learning and deep learning techniques. The objective of this paper is to provide a comprehensive overview of modern approaches to detecting anomalies in EHRs, outlining their strengths, limitations, and relevance to key healthcare challenges. We review traditional statistical methods alongside newer ML- and DL-based strategies and hybrid models, with particular attention to how these techniques support transparency and build clinical trust. Methods This paper presents a thorough and critical survey through systematic review (PRISMA-based) of the latest anomaly detection strategies in time-sequence data domains within electronic health record systems. Results We explore a broad spectrum of methodologies, including statistical models, supervised and unsupervised learning approaches, hybrid frameworks, and state-of-the-art ML-based techniques that collectively advance the precision and scalability of detecting anomalies in complex clinical datasets. In addition to mapping current capabilities, we address the enduring challenges that hinder widespread implementation and provide a forward-looking perspective on the future of anomaly detection in the data-rich landscape of modern healthcare. Summary The advancement in AI-based approaches is reported along with the basic principles of the individual approaches and their applicability. The increased availability of high-quality data, advancements in DL approaches, and enhanced computation power are leading to more frequent adaptation of DL-based approaches. Emerging DL-based approaches that have been adapted in other domains or recently applied in the EHR domain are also discussed in detail. Although DL-based approaches can improve model predictions by incorporating comorbidities, their application is limited in low-frequency data domains (e.g., when the total available data remains in the single digits). Therefore, the user must carefully consider the application based on data availability.

Anomaly detection↗

Xanthos-Lake Dataset

The Xanthos-Lake v1.0 dataset provides the input data, trained machine-learning models, and simulation outputs needed to characterize lake water balance, snow and ice conditions, and mixing-layer temperature within the Xanthos global hydrological modeling framework. The dataset supports lake representation across a wide range of lake sizes and hydroclimatic conditions by combining xLSIM, a basin-specific machine-learning emulator of lake snow, ice, ice-cover fraction, and mixing-layer temperature, with the Xanthos-Lake water-balance model. The archive contains NetCDF datasets used to train and evaluate xLSIM, trained model weights, processed meteorological and lake-property inputs, and basin- and lake-category-specific simulation outputs. These materials are organized into four primary data groups, described below. Snowice_model_inputs: Contains the NetCDF input data used to train xLSIM. The xLSIM machine-learning framework uses three lake-based datasets. The meteorological forcing dataset provides monthly relative humidity, specific humidity, surface wind speed, maximum and minimum air temperature, downward longwave and shortwave radiation, snowfall, surface air pressure, and total precipitation. Lake surface area is included as an additional static predictor. The target-state dataset provides lake ice thickness, snow depth, snow cover, and lake mixing-layer temperature, while a companion lake-surface dataset provides the lake ice-cover fraction. Before training, ice thickness and snow depth are converted from meters to centimeters, mixing-layer temperature is converted from kelvin to degrees Celsius and constrained to nonnegative values, and ice-cover fraction is converted from a fraction to a percentage. The predictor variables are normalized using statistics calculated across the selected lakes and time steps. Snowice_model_outputs: Contains the NetCDF outputs generated by xLSIM. For each basin, xLSIM produces a file containing observed and predicted lake-state variables for the training, validation, and testing periods. The modeled variables include lake ice thickness, snow depth, snow cover, mixing-layer temperature, and lake ice-cover fraction. For basins without a sufficiently persistent snow-and-ice signal, the emulator predicts only mixing-layer temperature. The outputs also include training and validation loss histories, the selected model configuration, identifiers of the lakes used in training, and SHAP-based feature-importance information at the global, lake, and seasonal-regime levels. The trained machine-learning model weights are provided separately within the dataset archive. Together, these files support model evaluation and subsequent coupling with the Xanthos-Lake water-balance framework. XanthosLAKES: Contains the NetCDF input data used by the Xanthos-Lake framework. Monthly meteorological inputs include relative and specific humidity, downward shortwave and longwave radiation, mean, maximum, and minimum air temperature, wind speed, precipitation, snowfall, and surface air pressure. Static lake-property datasets provide lake identifiers, geographic locations, surface area, volume, mean depth, elevation, drainage area, fetch, outlet-routing information, and associated Xanthos grid-cell attributes. Separate bathymetric datasets provide the coefficients of the area–depth and volume–depth relationships for each aggregated lake unit. GLEV-based records provide observed lake surface area and evaporation data used to initialize lake states, define reference conditions, and calibrate and evaluate the model. Xanthos-Lake Outputs: Contains the basin- and lake-category-specific NetCDF outputs generated by Xanthos-Lake. Monthly variables include lake surface area, storage volume, outlet discharge, evaporation rate, evaporation volume, lake–groundwater exchange, lake inflow, ice thickness, snow depth, snow-cover fraction, ice-cover fraction, and mixing-layer temperature. The files also contain lake-specific calibration and validation statistics, including normalized root-mean-square error, mean absolute error, Nash–Sutcliffe efficiency, Kling–Gupta efficiency, and percent bias. Stored calibrated and derived parameters include the weir discharge coefficient, fractional freeboard, groundwater exchange coefficient, reference water level, corresponding reference surface area and storage volume, weir-width adjustment factor, and the fraction of routed inflow entering the lake. Basin identifiers, lake category, simulation period, calibration and validation periods, and parameter-schema information are retained as NetCDF metadata.

Abeshu, Guta [Pacific Northwest National Laborator↗

Machine Learning Accelerates Innovation in Perovskite Manufacturing Scale-up (Final Technical Report (FTR))

We propose to address the challenge of the vast parameter space associated with perovskite manufacturing optimization, by developing a machine learning (ML)-assisted optimization framework for a scalable perovskite PV manufacturing tool. This framework will be interpretable, sequential, and rapidly adaptable to upgraded systems (e.g., via transfer learning). The tool is an open-air rapid spray plasma process (RSPP) of perovskite films, which has already been established at Stanford and is a unique platform to test and deploy the proposed ML-guided framework because the RSPP technique is able to conduct optimization experiments with a high throughput, and easily adjust a wide range of process variables.

14 SOLAR ENERGY↗

Actinium–DOTA coordination in water from hybrid ML/MM: Structure, free energies, and water-exchange pathways

Quantitative simulation of trivalent ƒ-block chelates in water remains challenging because bonded and non-bonded force-field models make different approximations for coordination structure, exchange dynamics, and ion–ligand interactions in highly charged systems. Here, we develop a hybrid machine-learning/molecular-mechanics (ML/MM) framework for Ac 3+ –DOTA in explicit solvent by training an E(3)-equivariant neural network potential (MACELES) on mechanically embedded QM/MM data for Ac aquo and Ac–DOTA species and coupling it to NAMD 2.14 with particle-mesh Ewald electrostatics. Nanosecond ML/MM trajectories remain numerically stable and preserve chelate integrity, yielding a compact DOTA inner shell with an inner-sphere water coordination number of CN Ac,O w ≈ 1.7 arising from a dynamic equilibrium between one- and two-water states (37.5% and 59.9% of frames; three waters 2.5%). A 5 ns potential of mean force shows two low-lying basins at CN Ac,O w ≈ 1 and CN Ac,O w ≈ 2. DFT end-state free energies are consistent with the ML/MM profile, and DFT minimum-energy paths provide a qualitative electronic-structure reference for the observed basin connectivity. State-resolved kinetics reveal picosecond water-exchange pathways that couple hydration changes to transient DOTA arm fluctuations, and training-set comparisons show that temperature-matched Ac–DOTA data optimize energy/force accuracy while more diverse solvated data improve charge prediction. Overall, the present hybrid ML/MM model provides a practical description of Ac 3+ –DOTA hydration thermodynamics and short-time exchange behavior in explicit water at MD-like cost.

Actinium↗

Toward equitable environmental exposure modeling through convergence of data, open, and citizen sciences: an example of air pollution exposure modeling amidst increasing wildfire smoke

Exposure modeling is critical in environmental epidemiology and human health but may face challenges (e.g., skewed data, unequal error, context-insensitive validation, and computational demands). Modeling decisions reflect the intended use of the models and the values that modelers prioritize. We aimed to provide a conceptual framework and machine learning (ML) modeling protocols that address these issues. With 500m-gridded hourly PM 2.5 and O 3 levels in Illinois before, during, and after the 2023 Canadian wildfire season as a motivating example, we conducted modeling experiments to evaluate modeling methods, guided by three domains we propose based on theories of science: 1) Data Diversity, leveraging open and citizen science data to enhance inclusivity, parsimony, and representativeness; 2) Equitable Accuracy, ensuring fairly distributed uncertainties across subpopulations; and 3) Sustainable Modeling, balancing accuracy with reducing computational demands to promote accessibility for under-resourced researchers. Here, we found that ML with publicly available data can achieve high accuracy. Depending on methods, performance may vary substantially, even with identical input data. Large but skewed data may reduce performance. Misuse of cross-validation protocols can underestimate prediction error; although we observed R 2 s of ∼98 %, the modeled estimates varied significantly, indicating the need for careful model validation. By using new modeling protocols including representativeness-considered training and validation data and a new loss function, we achieved high agreement between estimates and ground-based measurements (e.g., R 2 = ∼90 % for PM 2.5 ; ∼80 % for O 3 ), equally distributed errors across sociodemographic strata and urban–rural divides, and reduction in computation time—from several weeks or months to a few days.

Exposure assessment↗

AI-assisted object condensation clustering for calorimeter shower reconstruction at CLAS12

Several nuclear physics studies using the CLAS12 detector rely on the accurate reconstruction of neutrons and photons from its forward angle calorimeter system. These studies often place restrictive cuts when measuring neutral particles due to an overabundance of false clusters created by the existing calorimeter reconstruction software. In this work, we present a new AI approach to clustering CLAS12 calorimeter hits based on the object condensation framework. The model learns a latent representation of the full detector topology using GravNet layers, serving as the positional encoding for an event’s calorimeter hits which are processed by a Transformer encoder. This unique structure allows the model to contextualize local and long range information, improving its performance. Evaluated on one million simulated $e^-$ $+$ $p$ collision events, our method significantly improves cluster trustworthiness: the fraction of reliable neutron clusters, increasing from 8.88% to 30.73%, and photon clusters, increasing from 51.07% to 64.73%. In conclusion, our study also marks the first application of AI clustering techniques for hodoscopic detectors, showing potential for usage in many other experiments.

Calorimeters↗

Search for Stable and Low-Energy Ce–Co–Cu Ternary Compounds Using Machine Learning

Cerium-based intermetallics have garnered significant research attention as potential new permanent magnets. In this study, we explore the compositional and structural landscape of Ce−Co−Cu ternary compounds using a machine learning (ML)- guided framework integrated with first-principles calculations. We employ a crystal graph convolutional neural network (CGCNN), which enables efficient screening for promising candidates, significantly accelerating the material discovery process. With this approach, we predict five stable compounds, Ce 3 Co 3 Cu, CeCoCu 2 , Ce 12 Co 7 Cu, Ce 11 Co 9 Cu, and Ce 10 Co 11 Cu 4 , with formation energies below the convex hull, along with hundreds of low-energy (possibly metastable) Ce−Co−Cu ternary compounds. Firstprinciples calculations reveal that several structures are both energetically and dynamically stable. Notably, two Co-rich low-energy compounds, Ce 4 Co 33 Cu and Ce 4 Co 31 Cu 3 , are predicted to have high magnetizations.

Chemical structure↗

Critically assessing sodium-ion technology roadmaps and scenarios for techno-economic competitiveness against lithium-ion batteries

Sodium-ion batteries have garnered notable attention as a potentially low-cost alternative to lithium-ion batteries, which have experienced supply shortages and price volatility for key minerals. Here we assess their techno-economic competitiveness against incumbent lithium-ion batteries using a modelling framework incorporating componential learning curves constrained by minerals prices and engineering design floors. We compare projected sodium-ion and lithium-ion price trends across over 6,000 scenarios while varying Na-ion technology development roadmaps, supply chain scenarios, market penetration and learning rates. Assuming that substantial progress can be made along technology roadmaps via targeted research and development, we identify several sodium-ion pathways that might reach cost-competitiveness with low-cost lithium-ion variants in the 2030s. In addition, we show that timelines are highly sensitive to movements in critical minerals supply chains—namely that of lithium, graphite and nickel. Our modelled outcomes suggest that being price advantageous against low-cost lithium-ion variants in the near term is challenging and increasing sodium-ion energy densities to decrease materials intensity is among the most impactful ways to improve competitiveness.

25 ENERGY STORAGE↗

Single photon emitters in van der Waals solids for quantum photonics: materials, theory and molecular-scale characterization probes

Strong light–matter interactions in two-dimensional layered materials (2D materials) have attracted the interest of researchers from interdisciplinary fields for more than a decade now. A unique phenomenon in some 2D materials is their large exciton binding energies (BEs), increasing the likelihood of exciton survival at room temperature. It is this large BE that mediates the intense light–matter interactions of many of the 2D materials, particularly in their monolayer limit, where the interplay of excitonic phenomena poses a wealth of opportunities for high-performance optoelectronics and quantum photonics. Within quantum photonics, quantum information science (QIS) is growing rapidly, where photons are a promising platform for information processing due to their low-noise properties, excellent modal control, and long-distance propagation. A central element for QIS applications is a single photon emitter (SPE) source, where an ideal on-demand SPE emits exactly one photon at a time into a given spatiotemporal mode. Recently, 2D materials have shown practical appeal for QIS which is directly driven from their unique layered crystalline structure. This structural attribute of 2D materials facilitates their integration with optical elements more easily than the SPEs in conventional three-dimensional solid state materials, such as diamond and SiC. In this review article, we will discuss recent advances made with 2D materials towards their use as quantum emitters, where the SPE emission properties maybe modulated deterministically. Here, the use of unique scanning tunneling microscopy tools for the in-situ generation and characterization of defects is presented, along with theoretical first-principles frameworks and machine learning approaches to model the structure-property relationship of exciton–defect interactions within the lattice towards SPEs. Given the rapid progress made in this area, the SPEs in 2D materials are emerging as promising sources of nonclassical light emitters, well-poised to advance quantum photonics in the future.

2D layered materials↗

Machine learning accelerated prediction of Ce-based ternary compounds involving antagonistic pairs

The discovery of novel quantum materials within ternary phase spaces containing antagonistic pairs such as Fe with Bi, Pb, In, and Ag, presents significant challenges yet holds great potential. In this work, we investigate the stabilization of these immiscible pairs through the integration of Cerium (Ce), an abundant rare-earth and cost-effective element. By employing a machine learning (ML)-guided framework, particularly crystal graph convolutional neural networks (CGCNN), combined with first-principles calculations, we efficiently explore the composition/structure space and predict 9 stable and 37 metastable Ce-Fe-X (X=Bi, Pb, In, and Ag) ternary compounds. Our findings include the identification of multiple new stable and metastable phases, which are evaluated for their structural and energetic properties. These discoveries not only contribute to the advancement of quantum materials but also offer viable alternatives to critical rare earth elements, underscoring the importance of Ce-based intermetallic compounds in technological applications.

36 MATERIALS SCIENCE↗

Adaptive Reinforcement Learning (ARL) Control of a Multi-port Resonant Converter in UAV Systems

This study presents an adaptive reinforcement learning (ARL) control framework for a multi-port resonant converter used in hybrid unmanned aerial vehicle (UAV) power systems. The converter integrates high-frequency half-bridge input ports connected to a rectified engine–generator set and a battery energy storage system, along with a semi-bridgeless active rectifier supplying the propulsion load. A deep RL agent is trained to dynamically regulate inter-port phase-shift commands in real time based on flight conditions and load power demand. The ARL controller autonomously identifies phase-shift combinations that maximize conversion efficiency while maintaining stable and coordinated power flow, even under rapidly varying operating scenarios. This data-driven approach eliminates the need for explicit system modeling or extensive manual tuning and enables coordinated control among multiple power ports without inter-port communication. Experimental results validate that the ARL based strategy achieves reliable power sharing and consistently high-efficiency operation across diverse UAV operating conditions.

Asa, Erdem [ORNL] (ORCID:0000000190884812)↗

Cookie-Jar: An Adaptive Re-configurable Framework for Wireless Network Infrastructures

5G advancements like Massive Multiple Input Multiple Output (MIMO) bring high capacity and low latency, but also intensify interference challenges. Static and dynamic coordination techniques address this, often at the cost of increased power draw. We introduce Cookie-Jar (CJ), an interference coordination (IC) framework using reinforcement learning for multi-goal optimization. By dynamically adjusting network, power, and topology parameters based on real-time conditions, CJ improves Signal to Noise and Interference Ratio (SINR) while minimizing power consumption. Simulated 5G experiments showcase CJ's potential, achieving a 15% SINR improvement with near-identical power draw compared to existing methods.

Network↗

Intelligent Sampling of Extreme-Scale Turbulence Datasets for Accurate and Efficient Spatiotemporal Model Training

With the end of Moore’s law and Dennard scaling, efficient training increasingly requires rethinking data volume. Can we train better models with significantly less data via intelligent subsampling? To explore this, we develop SICKLE, a sparse intelligent curation framework for efficient learning, featuring a novel maximum entropy (MaxEnt) sampling approach, scalable training, and energy benchmarking. We compare MaxEnt with random and phase-space sampling on large direct numerical simulation (DNS) datasets of turbulence. Evaluating SICKLE at scale on Frontier, we show that subsampling as a preprocessing step can, in many cases, improve model accuracy and substantially lower energy consumption, with observed reductions of up to 38×.

Brewer, Wes [ORNL] (ORCID:0000000236393956)↗

Toward particle accelerator machine state embeddings as a modality for large language models

Understanding and diagnosing the state of a particle accelerator requires navigating high-dimensional control system data, often involving hundreds of interdependent parameters. We propose a novel multimodal embedding framework that jointly learns representations of machine states from both numerical control system readouts and natural language descriptions. This enables the translation of complex machine conditions into human-readable summaries while maintaining fidelity to the underlying physical system. The obtained embeddings are subsequently adapted to an open-weights large language model via cross-attention conditioning. We demonstrate a first implementation trained on European XFEL machine state data. This work covers the embedding model architecture, training methodology, and presents initial examples demonstrating the model's capabilities in action. Due to the general concept of machine state, the model can be easily adapted to other facilities and control system environments.

Accelerator Physics↗

Software Tools Ecosystem Project (STEP) Midyear Report CY2025

This document provides a technical project report for the first six months of 2025 for the Software Tools Ecosystem Project (STEP). The mission of STEP is to enable critical software tools to proactively adapt to emerging platform technologies (such as new accelerators, storage devices, network technologies, and smart devices) and emerging application use cases (such as advanced machine learning and workflow frameworks) so that they continue to meet the needs of scientific computing and provide a strong foundation for future Advanced Scientific Computing Research activities. Our challenges include the wide breadth of our stakeholders and rapidly evolving platform technology dependencies.

97 MATHEMATICS AND COMPUTING↗