Search NASASearch

SEARCH · Search NASA

Results for “Generative learning”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 181 records · Page 10

yProv4ML: Effortless provenance tracking for machine learning systems

The rapid growth in interest in deep learning and foundation models (FMs) in particular, has attracted the attention of a diverse range of researchers thanks to their generalization ability. However, the advent of these techniques has also brought to light the lack of transparency and rigor in the way development is pursued. In particular, the inability to determine the number of epochs and other hyperparameters in advance presents challenges in identifying the best model. To address this challenge, machine learning frameworks such as MLFlow can automate the collection of this type of information. However, these tools capture data using proprietary formats and pose little attention to lineage. This paper proposes yProv4ML, a framework that captures provenance information generated during machine learning processes in PROV-JSON format, with minimal code modification.

Machine learning

Superstructure Optimization of Waste Plastic Pyrolysis, Integrating Thermal, Catalytic, and Plasma Technologies with Machine Learning

Global plastic waste generation exceeds 430 million tonnes per year, yet fewer than 9% are recycled in the United States. Pyrolysis offers a chemical recycling route at scale, but existing techno-economic and life cycle assessments fix product yields to single pure polymers, producing economic and environmental outputs that break down when the feed composition changes. Here, we present a superstructure optimization framework that addresses this by embedding a composition-aware random forest yield predictor, trained on 566 pyrolysis experiments, within a full-scale process simulation. Product distributions update automatically as feed allocation shifts across four reactor chemistries: conventional thermal, catalytic (HZSM-5), thermal oxo-degradation, and nonequilibrium CO2 plasma. The optimal superstructure achieves minimum selling prices of −0.56 to −0.76/kg feed and global warming potentials of −0.276 to −0.322 kg CO2-eq/kg feed across four commodity price scenarios, confirming profitable, carbon-negative operation without tipping fees. Carbon abatement costs of $\$$0.46 to $\$$1.25/kg CO2-eq are competitive with direct air capture. Sensitivity analysis shows that the catalytic-plasma split fraction is the single largest driver of both economic and climate performance, while hydrocracking allocation in the wax upgrading stage is emission-neutral across the full variable range. Mixed plastic waste streams, evaluated as composition-variable feedstocks rather than pure resins, are profitable and carbon-negative across realistic market conditions. These results give a quantitative basis for reactor selection, circular economy investment, and policy design targeting chemical recycling on a large scale.

Life cycle assessment

Evaluation of Machine Learning Models for Automated Data Analysis in In-Service Nuclear Power Plant Inspections

The commercial nuclear power industry is facing a potential shortage of certified nondestructive evaluation (NDE) analysts to meet future in-service inspection demands. Automated data analysis (ADA) currently supports human inspectors in tasks such as eddy current evaluations for steam generator examinations. Machine learning (ML) systems are nearing the capability to pass performance demonstration tests for ultrasonic testing (UT) inspections of reactor pressure vessel upper head penetrations in nuclear power plants (NPPs). Current research and development is focused on assisted analysis (AA) of ADA versus fully automated examinations. This presentation will cover assessment of ML flaw detection on dissimilar metal weld (DMW) piping joints.

36 MATERIALS SCIENCE

Space Environments and Spacecraft Effects Concept: Transitioning Research to Operations and Applications

The National Aeronautics and Space Administration (NASA) is embarking on a course to expand human presence beyond Low Earth Orbit (LEO) while expanding its mission to explore the solar system. Destinations such as Near Earth Asteroids (NEA), Mars and its moons, and the outer planets are but a few of the mission targets. NASA has established numerous organizations specializing in specific space environments disciplines that will serve to enable these missions. To complement these existing discipline organizations, a concept is presented focusing on the development of a space environment and spacecraft effects organization. This includes space climate, space weather, natural and induced space environments, and effects on spacecraft materials and systems. This space environment and spacecraft effects organization would be comprised of Technical Working Groups (TWG) focusing on, for example: a) Charged Particles (CP), b) Space Environmental Effects (SEE), and c) Interplanetary and Extraterrestrial Environments (IEE). These technical working groups will generate products and provide knowledge supporting four functional areas: design environments, environment effects, operational support, and programmatic support. The four functional areas align with phases in the program mission lifecycle and are briefly described below. Design environments are used primarily in the mission concept and design phases of a program. Environment effects focuses on the material, component, sub-system and system-level selection and the testing to verify design and operational performance. Operational support provides products based on real time or near real time space weather observations to mission operators to aid in real time and near-term decision-making. The programmatic support function maintains an interface with the numerous programs within NASA and other federal agencies to ensure that communications are well established and the needs of the programs are being met. The programmatic support function also includes working in coordination with the program in anomaly resolution and generation of lesson learned documentation. The goal of this space environment and spacecraft effects organization is to develop decision-making tools and engineering products to support the mission phases of mission concept through operations by focusing on transitioning research to application. Products generated by this space environments and spacecraft effects organization are suitable for use in anomaly investigations. This paper will describe the organizational structure for this space environments and spacecraft effects organization, and outline the scope of conceptual TWG's and their relationship to the functional areas.

Edwards, D. L.

Space Environments and Spacecraft Effects Organization Concept

The National Aeronautics and Space Administration (NASA) is embarking on a course to expand human presence beyond Low Earth Orbit (LEO) while also expanding its mission to explore the solar system. Destinations such as Near Earth Asteroids (NEA), Mars and its moons, and the outer planets are but a few of the mission targets. Each new destination presents an opportunity to increase our knowledge of the solar system and the unique environments for each mission target. NASA has multiple technical and science discipline areas specializing in specific space environments disciplines that will help serve to enable these missions. To complement these existing discipline areas, a concept is presented focusing on the development of a space environments and spacecraft effects (SENSE) organization. This SENSE organization includes disciplines such as space climate, space weather, natural and induced space environments, effects on spacecraft materials and systems and the transition of research information into application. This space environment and spacecraft effects organization will be composed of Technical Working Groups (TWG). These technical working groups will survey customers and users, generate products, and provide knowledge supporting four functional areas: design environments, engineering effects, operational support, and programmatic support. The four functional areas align with phases in the program mission lifecycle and are briefly described below. Design environments are used primarily in the mission concept and design phases of a program. Engineering effects focuses on the material, component, sub-system and system-level selection and the testing to verify design and operational performance. Operational support provides products based on real time or near real time space weather to mission operators to aid in real time and near-term decision-making. The programmatic support function maintains an interface with the numerous programs within NASA, other federal government agencies, and the commercial sector to ensure that communications are well established and the needs of the programs are being met. The programmatic support function also includes working in coordination with the program in anomaly resolution and generation of lessons learned documentation. The goal of this space environment and spacecraft effects organization is to develop decision-making tools and engineering products to support all mission phases from mission concept through operations by focusing on transitioning research to application. Products generated by this space environments and effects application are suitable for use in anomaly investigations. This paper will describe the scope of the TWGs and their relationship to the functional areas, and discuss an organizational structure for this space environments and spacecraft effects organization.

Edwards, David L.

An Overview of the Space Environments and Spacecraft Effects Organization Concept

The National Aeronautics and Space Administration (NASA) is embarking on a course to expand human presence beyond Low Earth Orbit (LEO) while also expanding its mission to explore our Earth, and the solar system. Destinations such as Near Earth Asteroids (NEA), Mars and its moons, and the outer planets are but a few of the mission targets. Each new destination presents an opportunity to increase our knowledge on the solar system and the unique environments for each mission target. NASA has multiple technical and science discipline areas specializing in specific space environments fields that will serve to enable these missions. To complement these existing discipline areas, a concept is presented focusing on the development of a space environment and spacecraft effects (SESE) organization. This SESE organization includes disciplines such as space climate, space weather, natural and induced space environments, effects on spacecraft materials and systems, and the transition of research information into application. This space environment and spacecraft effects organization will be composed of Technical Working Groups (TWG). These technical working groups will survey customers and users, generate products, and provide knowledge supporting four functional areas: design environments, engineering effects, operational support, and programmatic support. The four functional areas align with phases in the program mission lifecycle and are briefly described below. Design environments are used primarily in the mission concept and design phases of a program. Environment effects focuses on the material, component, sub-system, and system-level response to the space environment and include the selection and testing to verify design and operational performance. Operational support provides products based on real time or near real time space weather to mission operators to aid in real time and near-term decision-making. The programmatic support function maintains an interface with the numerous programs within NASA, other federal government agencies, and the commercial sector to ensure that communications are well established and the needs of the programs are being met. The programmatic support function also includes working in coordination with the program in anomaly resolution and generation of lessons learned documentation. The goal of this space environment and spacecraft effects organization is to develop decision-making tools and engineering products to support all mission phases from mission concept through operations by focusing on transitioning research to application. Products generated by this space environments and effects application are suitable for use in anomaly investigations. This paper will describe the scope and purpose of the space environments and spacecraft effects organization and describe the TWG's and their relationship to the functional areas.

Edwards, David L.

Data Augmentation for Intelligent Contingency Management Using Generative Adversarial Neural Networks

Artificial intelligence (AI)-based techniques for intelligent contingency management (ICM) require that intelligent agents learn various aspects of system dynamics to create and execute contingencies. For high assurance contingency management, agents achieve the most compelling results through supervised or semi-supervised machine learning, for which agents require large datasets to learn the dynamics of the system. Unfortunately, data collection in aerospace applications can be costly, due to both time and resources. Presented work describes a framework for data augmentation of ICM databases containing training data for machine learning models. This framework populates the database with the outputs of generative adversarial network (GAN) models that were trained on flight data. Methods for evaluating the suitability of these models based on the equations of motion, as well as other physical constraints, are discussed. The paper demonstrates the utility of this database for training intelligent agents on the NASA T2 generic transport aircraft model and experimental vertical takeoff and landing (VTOL) simulation model.

Generative Machine Learning

Data Augmentation for Intelligent Contingency Management Using Generative Adversarial Neural Networks

Artificial intelligence (AI)-based techniques for intelligent contingency management (ICM) require that intelligent agents learn various aspects of system dynamics to create and execute contingencies. For high assurance contingency management, agents achieve the most compelling results through supervised or semi-supervised machine learning, for which agents require large datasets to learn the dynamics of the system. Unfortunately, data collection in aerospace applications can be costly, due to both time and resources. Presented work describes a framework for data augmentation of ICM databases containing training data for machine learning models. This framework populates the database with the outputs of generative adversarial network (GAN) models that were trained on flight data. Methods for evaluating the suitability of these models based on the equations of motion, as well as other physical constraints, are discussed. The paper demonstrates the utility of this database for training intelligent agents on the NASA T2 generic transport aircraft model and experimental vertical takeoff and landing (VTOL) simulation model.

Generative Machine Learning

Generation of Continental Scale Percent Tree Cover Product Using Deep-learning and Multi-scale Remote Sensing Data

Spatially explicit percent tree cover (TC) estimation is critical for mapping forest aboveground biomass and its dynamics. While various TC products have been developed, there has not been a generalized framework that can be applied to diverse terrestrial ecosystems due to underlain extreme complexities. Deep learning algorithms can learn a spatial pattern and radiometric characteristics of tree canopy as a robust approximation of physical or empirical models, and thus have emerged as promising and efficient tools for large-scale TC mapping. In this study, we synergistically use very high-resolution aerial imageries (National Agriculture Imagery Program, NAIP) and medium resolution Landsat data to map continental-scale TC (CONUS and Mexico) through a hierarchical deep learning approach (Convolutional Neural Network), i.e., NAIP TC generated from a NAIP model is utilized to train a Landsat model. The produced TC product (hereafter, NEX-TC) is able to capture the spatial pattern of TC distribution and its changes driven by natural disturbance and human land management. We further explore and analyze the reliability and potential uncertainty of the NEX-TC by comparing it to lidar- (lidar-TC), National Land Cover Database (NLCD-TC), and MODIS Vegetation Continuous Field (MODIS-TC). This evaluation practice reveals that TC products based on passive optical sensors tend to underestimate TC across all land cover types while Landsat-based TCs (i.e., NEX-TC & NLCD-TC) perform better than the coarser MODIS TC estimate. Our results show that the NEX-TC is generally comparable to NLCD-TC but it particularly outperforms NLCD-TC and MODIS-TC over the dense forests where lidar-TC indicates >80% TC. These results indicate that our hierarchical deep learning approach and TC product will be effective and useful for characterizing large-scale tree cover and possibly associated carbon dynamics.

Landsat

Large-Scale High-Resolution Coastal Mangrove Forests Mapping Across West Africa With Machine Learning Ensemble and Satellite Big Data

Coastal mangrove forests provide important ecosystem goods and services, including carbon sequestration, biodiversity conservation, and hazard mitigation. However, they are being destroyed at an alarming rate by human activities. To characterize mangrove forest changes, evaluate their impacts, and support relevant protection and restoration decision making, accurate and up-to-date mangrove extent mapping at large spatial scales is essential. Available large-scale mangrove extent data products use a single machine learning method commonly with 30 m Landsat imagery, and significant inconsistencies remain among these data products. With huge amounts of satellite data involved and the heterogeneity of land surface characteristics across large geographic areas, finding the most suitable method for large-scale high-resolution mangrove mapping is a challenge. The objective of this study is to evaluate the performance of a machine learning ensemble for mangrove forest mapping at 20 m spatial resolution across West Africa using Sentinel-2 (optical) and Sentinel-1 (radar) imagery. The machine learning ensemble integrates three commonly used machine learning methods in land cover and land use mapping, including Random Forest (RF), Gradient Boosting Machine (GBM), and Neural Network (NN). The cloud-based big geospatial data processing platform Google Earth Engine (GEE) was used for pre-processing Sentinel-2 and Sentinel-1 data. Extensive validation has demonstrated that the machine learning ensemble can generate mangrove extent maps at high accuracies for all study regions in West Africa (92%–99% Producer’s Accuracy, 98%–100% User’s Accuracy, 95%–99% Overall Accuracy). This is the first-time that mangrove extent has been mapped at a 20 m spatial resolution across West Africa. The machine learning ensemble has the potential to be applied to other regions of the world and is therefore capable of producing high-resolution mangrove extent maps at global scales periodically.

coastal environment

Predictive Chemical Kinetic Modeling: Where We Succeed, Where We Struggle, and What Comes Next

Chemical kinetic modeling plays a foundational role in fields ranging from energy to environmental science, pharmaceuticals, and advanced materials. The past two decades have seen remarkable progress, particularly in modeling gas-phase reactions for thermochemical processes, leading to impactful industrial applications such as steam cracking and air quality management. However, new challenges are emerging. The successful development of systematic methodologies for the description of gas-phase kinetics opens the possibility to apply the same approach to the study of more challenging systems. Here, we review recent advances, including ab initio transition state theory-based master equation estimation of elementary rates, automated mechanism generation, machine-learning-assisted kinetics, and uncertainty quantification, and discuss the advances needed to apply the same methodological approach in areas such as heterogeneous catalysis, electrochemistry, liquid-phase and solid-state reactivity, and multiscale model integration. We advocate for the development of targeted tools, especially methods that go beyond empirical tuning toward first-principles-based predictions. We highlight the need for accessible software and AIaugmented workflows to democratize modeling for industry and academia alike. In this perspective, we call attention to not only what has worked but also what remains unsolved, advocating to avoid overemphasizing successes in scientific works at the expense of realism. The next decade should focus on predictive capability, physical accuracy, and community infrastructure (e.g., databases and services) to enable innovation across diverse fields. We argue that kinetic modeling, properly equipped, can accelerate discovery far beyond its traditional domains.

ab initio calculations

Enhancing Unknown Waveform Detection by Learning Intra and Inter-domain Dependencies with Advanced Attention Fusion Mechanisms

Detection of unknown waveforms in mission-critical communications is a crucial area of interest for the Department of Energy (DoE). Traditional methods and recent deep learning-based approaches often assume that the training set includes all possible classes, which is impractical for detecting new waveforms. This limitation gives rise to the problem of open-set recognition (OSR), which involves correctly identifying known classes while detecting and rejecting unknown or unseen classes. To address this limitation, we propose a novel dual-domain complex-valued neural architecture that jointly processes time-domain and frequency-domain signal representations using transformer mechanisms. A transformer model is a deep learning architecture that uses self-attention mechanisms to process and learn relationships in sequential data. Our model employs a cosine similarity loss to extract domain-specific features and incorporates a transformer architecture in the latent space to weigh the importance of different features from the time and frequency domains. The transformer layer includes stacked self-attention and cross-attention modules to learn intra-domain and inter-domain dependencies, creating a more holistic signal representation. An attention-based fusion module intelligently combines the time and frequency-domain features using multi-head attention, enabling the network to learn the optimal feature for each domain in each input signal. Quantitative results demonstrate the impact of these architectural choices on overall performance, showing significant improvement after incorporating self and cross-attention modules and using complex attention fusion over simple weighted fusion. Our ongoing work will focus on addressing the limitations of threshold-based OSR methods by developing a novel generative framework that integrates a conditional diffusion probabilistic model (DPM). DPM is a generative framework that learns to synthesize complex data by reversing a gradual noising process using a neural network trained to denoise step-by-step. Our goal is to leverage the inherent strengths of DPMs for identifying unknown signals more robustly. One primary advantage of using a DPM is its ability to provide a more reliable anomaly score based on the model's reconstruction error, rather than relying solely on classifier confidence. Additionally, the iterative denoising process of DPMs makes this approach naturally resilient to low Signal-to-Noise Ratio (SNR) conditions, where traditional methods often fail. By implementing this generative framework, we aim to enhance the model's capability to accurately detect unknown waveforms and maintain performance in challenging environments.

99 - GENERAL AND MISCELLANEOUS

Generative Electrolyte Solvent and Formulation Discovery

Molecular mixtures and/or formulations are of great importance in fields ranging from materials science to pharmaceuticals to chemistry. In batteries, electrolytes are complex molecular mixtures consisting of multiple salts and solvents and additives at different concentrations that dictate battery capacity, safety, and cycle life, among others. Unfortunately, due to the complex composition and infinite design space as well as the conflicting property requirements, electrolyte design is the rate-determining step in the design of next generation battery chemistries. In this work, we develop a transformer-based generative AI model − ElectrolyteGPT − capable of generating solvents and electrolyte formulations to satisfy a wide range of desired property requirements. First, we curate an electrolyte-relevant database and develop a new line notation for formulations. Then, we show that ElectrolyteGPT can generate solvents and formulations conditioned on a wide range of important electrolyte properties such as ionic conductivity, oxidative stability, Coulombic efficiency, viscosity, and more. Finally, we experimentally synthesize the generated solvents and fabricate the electrolyte formulations and show that they can meet the desired property requirements and enable longterm cycling in energy-dense anode-free lithium metal batteries. Our work showcases the ability of generative models to address challenges in molecular mixture design for next generation batteries.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Custom-trained Machine-learning Interatomic Potentials: ZnCl2 Aqueous Solution

This dataset was generated using an iterative active-learning strategy implemented in the ArcaNN software package (https://github.com/arcann-chem/arcann_training) to train machine-learning interatomic potentials for aqueous ZnCl2 solutions. Each active-learning cycle consisted of three stages: training, exploration, and labeling. The initial training set combined configurations generated in this work from enhanced-sampling ab initio molecular dynamics simulations with configurations from a previously reported neural-network-potential study of aqueous ZnCl2. The enhanced-sampling ab initio molecular dynamics simulations involved Zn–Cl separation and the chloride coordination number around Zn²? as collective variables. These configurations served as the seed dataset. Subsequent active-learning cycles expanded the training set by identifying and labeling configurations that were poorly represented by the current models, thereby improving coverage of ion-association states and changes in local coordination and charge-state environments relevant to the solution free-energy landscape. For all selected configurations, single-point calculations of the total energies and atomic forces were performed within density functional theory using the CP2K Quickstep module. Reference calculations employed the revPBE-D3 and r2SCAN exchange-correlation functionals. Motivated by recent work on aqueous Zn²?, the main revPBE calculations omitted D3 dispersion contributions involving Zn²?, while retaining the D3 correction for water and chloride. For comparison, fully dispersion-corrected revPBE-D3 reference calculations were also performed, with D3 applied to all species, including Zn²?. Valence electrons were treated explicitly, while core electrons were represented using norm-conserving Goedecker–Teter–Hutter pseudopotentials. The wave functions were expanded using the mixed Gaussian-and-plane-wave scheme with TZV2P-MOLOPT basis sets for all elements and a 600 Ry auxiliary plane-wave cutoff for the electron density. Self-consistent-field convergence was accelerated using the orbital-transformation and Direct Inversion in the Iterative Subspace algorithms, with a convergence threshold of 10?6. All single-point calculations were performed in periodic orthorhombic cells. The CELL_REF keyword in CP2K was used to define a fixed reference cell with a box length of 25 Å. This treatment ensured a consistent reference for configurations extracted from NpT trajectories with fluctuating cell dimensions. The resulting DFT energies and atomic forces constitute the ground-truth labels used to train the MLIPs. The resulting MLIP was trained for aqueous ZnCl2 solutions spanning concentrations from 0 to 30 molal and a broad pH range, from strongly acidic to strongly basic conditions. Representative examples of configurations included in the MLIP training dataset are provided below. These include 1) Representative configurations from the dataset labeled at the revPBE-D3 level, with D3 dispersion interactions involving Zn2+ excluded (revPBE-wo-D3). 2) Representative configurations from the dataset labeled at the fully dispersion-corrected revPBE-D3 level, with D3 interactions applied to all species, including Zn2+ (revPBE-D3). 3) Representative configurations from the dataset labeled at the r2SCAN level of theory (r2SCAN).

Dinpajooh, Mohammadhasan [Pacific Northwest Nation

Reduced Erosion Augments Soil Carbon Storage Under Cover Crops

ABSTRACT Cover crops, a promising strategy to increase soil organic carbon (SOC) storage in croplands and mitigate climate change, have typically been shown to benefit soil carbon (C) storage from increased plant C inputs. However, input‐driven C benefits may be augmented by the reduction of C outputs induced by cover crops, a process that has been tested by individual studies but has not yet been synthesized. Here we quantified the impact of cover crops on organic C loss via soil erosion (SOC erosion) and revealed the geographical variability at the global scale. We analyzed the field data from 152 paired control and cover crop treatments from 57 published studies worldwide using meta‐analysis and machine learning. The meta‐analysis results showed that cover crops widely reduced SOC erosion by an average of 68% on an annual basis, while they increased SOC stock by 14% (0–15 cm). The absolute SOC erosion reduction ranged from 0 to 18.0 Mg C −1 ha −1 year −1 and showed no correlation with the SOC stock change that varied from −8.07 to 22.6 Mg C −1 ha −1 year −1 at 0–15 cm depth, indicating the latter more likely related to plant C inputs. The magnitude of SOC erosion reduction was dominantly determined by topographic slope. The global map generated by machine learning showed the relative effectiveness of SOC erosion reduction mainly occurred in temperate regions, including central Europe, central‐east China, and Southern South America. Our results highlight that cover crop‐induced erosion reduction can augment SOC stock to provide additive C benefits, especially in sloping and temperate croplands, for mitigating climate change.

Huang, Wenjuan [Department of Ecology, Evolution,

Datasets for Custom-trained Machine-learning Interatomic Potentials: Nitric Acid Aqueous Solution

This dataset was generated using an iterative active learning strategy with the ArcaNN software package (https://github.com/arcann-chem/arcann_training) to train machine-learning interatomic potentials (MLIPs) for aqueous nitric acid. Each active-learning cycle consisted of three stages: (1) training, (2) exploration, and (3) labeling. The initial training set comprised approximately 800 randomly selected configurations from a previous study by Lewis et al. (https://doi.org/10.1021/jp205510q), which investigated nitric acid solutions at 2, 3, 4, and 5 mol/L. For all configurations, single-point calculations of atomic forces and total energies were performed at the quantum density functional theory BLYP-D2 and PBE-D3 levels of theory using the CP2K Quickstep module. Valence electrons were treated explicitly, while core electrons on all atoms were represented by norm-conserving Goedecker–Teter–Hutter (GTH) pseudopotentials. Long-range dispersion interactions were accounted for using Grimme dispersion corrections. Wave functions were expanded in a mixed Gaussian-and-plane-wave scheme using TZV2P-MOLOPT basis sets for all elements and an 800 Ry auxiliary plane-wave cutoff for the electron density. Self-consistent field convergence was accelerated using orbital transformation and Direct Inversion in the Iterative Subspace, with a convergence threshold of 10^{-6}. All single-point calculations were carried out in periodic orthorhombic cells whose dimensions match those of the molecular configurations sampled from earlier trajectories. The CELL_REF keyword in CP2K was used to define a fixed reference cell, ensuring consistency in the reference data used for MLIP training, particularly when cell fluctuations are present in NpT simulations. The resulting high-fidelity energies and forces constitute the ground-truth labels used to train the MLIPs contained in this dataset.

Dinpajooh, Mohammadhasan [Pacific Northwest Nation

Solidification cracking of refractory alloys: a computational and machine learning study to investigate composition-dependence for improved weldability and additive manufacturability

Large-batch numerical, CALculation of PHAse Diagrams (CALPHAD)-based solidification cracking calculations are performed and then analyzed with machine learning methods to generate models that relate chemistry of refractory alloys to cracking susceptibility. Kou’s solidification cracking index is used to study the refractory alloys including O, N, C binary mixtures with Mo, Ta, Nb, and W, the molybdenum-based TZM, Niobium-based C103, and Tantalum-based T111 and Ta-10 W, as well as hypothetical refractory ternary alloys. Findings strongly validate Kou’s Crack Susceptibility Index (CSI) against Varestraint test data for Nb- and Ta-based alloys, establishing CSI thresholds where refractory alloys with CSI < 15,000 K are likely weldable, CSI > 15,000 K are prone to cracking, and CSI > 25,000 K are likely unweldable (or unprintable). Furthermore, interstitial elements C, N, and O significantly increase crack susceptibility, with some existing material specifications coinciding with peak cracking susceptibility concentrations. Finally, machine learning-derived elemental potency factors enable rapid prediction of CSI from alloy chemistry for C103, TZM, Ta-10 W, and T-111 alloys. These results provide practical guidance for feedstock selection, powder reuse limits, and alloy specification amendments for welding and additive manufacturing applications.

36 MATERIALS SCIENCE