Search NASASearch

SEARCH · Search NASA

Results for “Data Mining”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

331 records · Page 2

4D-STEM Mapping of Nanocrystal Reaction Dynamics and Heterogeneity in a Graphene Liquid Cell

Chemical reaction kinetics at the nanoscale are intertwined with heterogeneity in structure and composition. However, mapping such heterogeneity in a liquid environment is extremely challenging. Here, in this work, we integrate graphene liquid cell (GLC) transmission electron microscopy and four-dimensional scanning transmission electron microscopy to image the etching dynamics of gold nanorods in the reaction media. Critical to our experiment is the small liquid thickness in a GLC that allows the collection of high-quality electron diffraction patterns at low dose conditions. Machine learning-based data-mining of the diffraction patterns maps the three-dimensional nanocrystal orientation, groups spatial domains of various species in the GLC, and identifies newly generated nanocrystallites during reaction, offering a comprehensive understanding on the reaction mechanism inside a nanoenvironment. This work opens opportunities in probing the interplay of structural properties such as phase and strain with solution-phase reaction dynamics, which is important for applications in catalysis, energy storage, and self-assembly.

four-dimensional scanning transmission electron mi

Radiolytically reworked Archean organic matter in a habitable deep ancient high-temperature brine

Abstract Investigations of abiotic and biotic contributions to dissolved organic carbon (DOC) are required to constrain microbial habitability in continental subsurface fluids. Here we investigate a large (101–283 mg C/L) DOC pool in an ancient (>1Ga), high temperature (45–55 °C), low biomass (10 2 −10 4 cells/mL), and deep (3.2 km) brine from an uranium-enriched South African gold mine. Excitation-emission matrices (EEMs), negative electrospray ionization (–ESI) 21 tesla Fourier-transform ion cyclotron resonance mass spectrometry (FT-ICR MS), and amino acid analyses suggest the brine DOC is primarily radiolytically oxidized kerogen-rich shales or reefs, methane and ethane, with trace amounts of C 3 –C 6 hydrocarbons and organic sulfides. δ 2 H and δ 13 C of C 1 –C 3 hydrocarbons are consistent with abiotic origins. These findings suggest water-rock processes control redox and C cycling, helping support a meagre, slow biosphere over geologic time. A radiolytic-driven, habitable brine may signal similar settings are good targets in the search for life beyond Earth.

Science & Technology - Other Topics

Innovations Driven by Advanced Characterization to Strategize Critical Mineral Production and Beneficial Reuse from Fossil Energy Waste

Critical minerals (CM), such as rare earth elements (REE), cobalt, nickel, and lithium, have important uses in modern electronics and advanced manufacturing, yet are vulnerable to potential supply chain disruptions. Relatively abundant and readily available fossil energy (FE) wastes, such as coal combustion ash, acid mine drainage (AMD) and treatment solids (AMD solids), and Oil and Gas (O&G) drilling wastes (drill cuttings and produced waters) are under consideration as CM feedstocks. The National Energy Technology Laboratory (NETL) has studied CM resources for various FE wastes as part of the U.S. Department of Energy’s mission of bolstering the domestic CM supply, and makes the data available to the public on EDX at sites such as the NEWTS group. Advanced characterization utilizing synchrotron x-ray techniques coupled with laboratory extractions has been performed to identify CM hosting phases in these FE wastes to inform CM recoverability mechanisms. Novel methods to selectively recover CMs while co-producing other valuable byproducts have been developed. Successful examples discussed here include: (1) The identification of REE/Co/Ni/Sc binding and hosting phases in select FE waste (coal combustion ash and AMD solids), resulting in the development of a patented CM step-extraction process, (2) coupled production of functional sorbents from these extraction wastes and for CM recovery. A pilot-scale testing to evaluate the patent’s technical feasibility for extracting REE from coal ash on a barrel scale has been successfully performed. Additionally, (3) evaluation and measurements of brine geochemistry from U.S. O&G produced waters has informed a high Li recovery potential from Marcellus Shale produced water. NETL researchers have been developing tailored pre-treatment processes, an innovative and highly durable lithium sorbent, and geochemical model guided precipitation to accelerate Li production from the Marcellus Shale produced waters. These innovations driven by characterization are integral for maximizing and advancing the potential for CM recovery while offsetting the cost and environmental footprint for FE waste management.

critical mineral processing

From Rules to Reasoning: A Survey of Large Language Model-Based Approaches to Scientific Hypothesis and Idea Generation

Scientific hypothesis generation represents a fundamental challenge in contemporary research due to exponentially expanding literature volumes and increasing disciplinary specialization. Large language models (LLMs) have emerged as transformative tools for automated scientific discovery, moving beyond traditional rule-based and literature-mining approaches. Four paradigmatic approaches define current LLM-driven hypothesis generation: direct prompting and fine-tuning methods, knowledge-enhanced frameworks integrating retrieval-augmented generation (RAG), multi-agent collaborative systems simulating research teams, and reasoning-focused approaches implementing cognitive architectures. Domain-specific applications demonstrate statistical equivalence to human expert performance in social psychology, experimental validation in biomedical research, and near-expert quality in astronomy. Evaluation methodologies encompass human expert assessment, LLM-as-judge frameworks, and comprehensive benchmarking systems. Technical challenges include hallucination management, knowledge integration limitations, and balancing novelty with feasibility. Future directions emphasize hybrid neural-symbolic architectures and sophisticated human-AI collaboration models for responsible scientific discovery acceleration.

AI-driven discovery

Boron-Based Neutron Scintillator Screen Characterization with X-Rays and Neutrons

Recent work on boron-based neutron scintillator screens suggests these screens can offer superior performance when compared to commonly used screens. Borated neutron scintillator screens perform well in terms of light output (5-6 times greater than a standard Gadox screen) and detection effi-ciency (larger than standard LiF+ZnS screens). However, previously manu-factured boron-based screens have exhibited non-uniform surface coating and a poor mixture between phosphor and converter particles. The objective of this work was to evaluate newly fabricated scintillator screens to deter-mine if enhanced fabrication methods produced a more homogeneous distribution between neutron converter and scintillation phosphor particles. Uniformity of scintillator material deposition was also inspected. This new iteration of screens appeared more uniform than previous generations with the new coating method improving surface chemistry and scintillator material homogeneity. Additionally, a new methodology for screen characterization, involving the correlation of a neutron image taken with a borated scintillator screen to X-ray computed tomography of that same screen, was demonstrated to elucidate a relationship between scintillator screen thickness and relative light output of the screen under neutron exposure. This method suggested that the ideal thickness of scintillator material was ~150 µm to maximize light output of the screen.

36 - MATERIALS SCIENCE

Internet of Things Data Characterization Process: Pattern of Life Behavioral Data Study

The HoneyBee™ TARDIS LDRD team completed a data scoping study that identified the initial processes and procedures to baseline the normal and expected behaviors during operability and interoperability of Internet of Things (IoT) device networks. This research is the initial step in developing a process (or methodology) to inform a much broader information framework incorporating machine learning to determine device pattern-of-life which enables the detection of abnormal IoT behaviors on an individual device, as well as in the context of a larger network.

97 MATHEMATICS AND COMPUTING

Circumventing data imbalance in magnetic ground state data for magnetic moment predictions

Abstract Magnetic materials play a crucial role in the transition to more sustainable forms of energy and electric vehicles. There is an anticipated shortage in magnetic materials in the future, and as a result there is an urgent need to discover and design new magnetic materials. Computational magnetic material design using density functional theory is daunting because of the challenge in identifying magnetic ground states from a combinatorially large set of possibilities. Machine learning offers a path forward by enabling efficient surrogate models that can more readily enumerate these states, but there is a dearth of training data available, and what is available tends to be imbalanced with too much non-magnetic data. In this work we show that the discrete and previously tackled data imbalance that exists at the level of the magnetic ordering leads to an imbalanced continuous distribution with many zeros when the data is unraveled at the atomic magnetic moment level, which subsequently leads to models with low accuracy for magnetic properties. We mitigate this by using a two-part model framework. Our scheme is able to classify atoms into magnetic and non-magnetic with an F1 score and Matthew’s correlation coefficient (MCC) of ~91% and then to provide an implicit embedding representation that maps directly onto the magnitude of the magnetic moment with a mean absolute error of 0.1 μ B . Beyond screening for new magnetic materials, we demonstrate an additional practical use case of our scheme: the provision of good initial guesses for magnetic moments in first-principles electronic relaxations. Such initialization is shown to lead to faster convergence to configurations that lie closer to the ground state.

Computer Science

Overview of IMPACT Data Acquisition System and Data Reduction Process

This report documents the development of the data acquisition system (DAS) and data reduction methodologies for the Irradiated Material Property Accelerated Characterization Test (IMPACT) experiment at the Advanced Test Reactor (ATR). The IMPACT experiment is designed to enable in-pile measurement of thermal conductivity in metallic nuclear fuels, specifically U-10Zr, using an instrumented thermal conductivity probe. The DAS supports both passive temperature monitoring and active thermal interrogation of the probe through controlled AC and DC excitation. Significant modifications to laboratory-scale systems were required to accommodate the higher resistance paths associated with the in-pile application. Custom electronics and relay-controlled measurement sequencing were developed to enable the measurement and sufficient power delivery to the sensing region. A reduced-order, axisymmetric thermal model based on the thermal quadrupoles method is presented to support data interpretation. This model enables efficient evaluation of transient heat transfer behavior and facilitates solution of the inverse problem required to extract thermal properties from measured signals. Multiple boundary condition formulations are discussed to address varying experimental time scales and geometries. Additionally, machine learning techniques are introduced to support data reduction and improve confidence in inverse solutions. Convolutional neural networks are applied to identify the presence of gas gaps and other evolving geometric features that significantly impact thermal response during irradiation. These efforts contribute to the broader integration of digital twin frameworks and real-time modeling capabilities within the Advanced Fuels Campaign.

46 - INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AN

Framework of compressive sensing and data compression for 4D-STEM

Four-dimensional Scanning Transmission Electron Microscopy (4D-STEM) is a powerful technique for high-resolution and high-precision materials characterization at multiple length scales, including the characterization of beam-sensitive materials. However, the field of view of 4D-STEM is relatively small, which in absence of live processing is limited by the data size required for storage. Furthermore, the rectilinear scan approach currently employed in 4D-STEM places a resolution- and signal-dependent dose limit for the study of beam sensitive materials. Improving 4D-STEM data and dose efficiency, by keeping the data size manageable while limiting the amount of electron dose, is thus critical for broader applications. Here we introduce a general method for reconstructing 4D-STEM data with subsampling in both real and reciprocal spaces at high fidelity. The approach is first tested on the subsampled datasets created from a full 4D-STEM dataset, and then demonstrated experimentally using random scan in real-space. The same reconstruction algorithm can also be used for compression of 4D-STEM datasets, leading to a large reduction (100 times or more) in data size, while retaining the fine features of 4D-STEM imaging, for crystalline samples.

4D-STEM

AI-Batt (Autonomous Identification of Battery Life Models) [SWR 21-36]

Autonomous Identification of Battery Life Models (AI-Batt) AI-Batt is a MATLAB code base for developing lifetime models for batteries from accelerated aging data. The code base provides many functions for processing, visualizing, and modeling battery aging data, making the data processing, exploration, and modeling workflow substantially faster. These tools are tailored for working with battery aging data sets, which usually consist of many separate time-series for each cell, with many test conditions and possible replicates at each condition, which makes it difficult to simply process or visualize the data set. Complex modeling tasks, such as cross-validation, sensitivity analysis, and uncertainty quantification have been implemented to enable thorough statistical investigation of model predictions. Additionally, several machine-learning algorithms are implemented to autonomously identify suitable models via symbolic regression. Data processing functions automatically cast data from the struct data type, which is commonly used to store experimental data, but is not an acceptable input for most algorithms, to the table data type, which can be easily used as input to any optimization algorithm. Also, the data can be separated into time-invariant and time-variant data tables, which is helpful for exploring the data set as well as developing separate models for time-variant and time-invariant aging mechanisms. For example, in aging tests with constant temperature, temperature is a time-invariant experimental condition. Visualization tools enable plotting of data, model fits, and model simulations possible with single-line function calls, empowering data exploration of complex data sets with both time-varying and time-invariant trends. Plots can be automatically generated for the whole data set, or separated by data group (groups of test replicates) or individual data series. Data points or data series can be automatically colored by the value of a variable with a variety of color maps, and model predictions can also be colored by the value of a fit statistic. Comparisons between data sets and the predictions/simulations of different models on the same data set can be easily plotted as well. Distributions of parameter values from bootstrap resampling can be plotted to visualize the reliability of parameter estimation, or determine any correlations between parameters. Modeling tools handle the complex task of creating and parsing symbolic equations for modeling battery lifetime. Equations are parsed to grab relevant data variables, parameter values, or specified sub-models for input into optimization, evaluation, or simulation functions. Models can be optimized locally (one set of parameters for each data series), bi-level (some parameters shared across the data set), or globally (single set of parameters for all data). Functions implementing symbolic regression algorithms help users to discover effective model equations, even in poorly sampled, high-dimensional data.

Smith, Kandler [National Renewable Energy Lab. (NR

Progress Report on SFR Metallic Fuel Data Qualification

This report summarizes the progress of SFR metallic fuel qualification related activities, which are focused on providing quality assurance relevant information applicable to experiments irradiated during the Integral Fast Reactor (IFR) program. An overview of the metallic fuel performance data and the associated databases, including the EBR-II Fuels Irradiation & Physics Database (FIPD), Out-of-Pile Transient Database (OPTD), and TREAT Experimental Relational Database (TREXR) is included. The legacy data in the databases, including as-built, post-irradiation examination (PIE), operating parameters, and out-of-pile experiment post-test data are introduced. The SFR metallic fuel Quality Assurance Program Plan (QAPP) and its implementation to qualify these legacy data is described in detail. Important PIE data QA documents and the specifications of seven types of PIE measurements (contact profilometry, laser profilometry, neutron radiography, gamma scan, fission gas release fission gas chemistry, and metallography) are provided. Examples of the implementation of the QAPP to qualify each of those types of PIE data are provided.

Mo, Kun [Argonne National Laboratory (ANL), Argonn