Search NASASearch

SEARCH · Search NASA

Results for “generative machine learning”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

Multi Modality Brain Mapping System (MBMS) Using Artificial Intelligence and Pattern Recognition

A Multimodality Brain Mapping System (MBMS), comprising one or more scopes (e.g., microscopes or endoscopes) coupled to one or more processors, wherein the one or more processors obtain training data from one or more first images and/or first data, wherein one or more abnormal regions and one or more normal regions are identified; receive a second image captured by one or more of the scopes at a later time than the one or more first images and/or first data and/or captured using a different imaging technique; and generate, using machine learning trained using the training data, one or more viewable indicators identifying one or abnormalities in the second image, wherein the one or more viewable indicators are generated in real time as the second image is formed. One or more of the scopes display the one or more viewable indicators on the second image.

Kateb, Babak

Reusing Data and Metadata to Create New Metadata Through Machine-Learning & Other Programmatic Methods

Recent improvements in natural language processing (NLP) enable metadata to be created programmatically from reused original metadata or even the dataset itself. Transfer-learning applied to NLP has greatly improved performance and reduced training data requirements. In this talk, we’ll compare machine-generated metadata to human-generated metadata and discuss characteristics of metadata and data archives that affect suitability for machine-learning reuse of metadata. Where as human-generated metadata is often populated once, populated from the perspective of data supplier, populated by many individuals with different words for the same thing, and limited in length, machine-generated metadata can be updated any number of times, generated from the perspective of any user, constrained to a standardized set of terms that can be evolved over time, and be any length required. Machine-learning generated metadata offers benefits but also additional needs in terms of version control, process transparency, human-computer interaction, and IT requirements. As a successful example, we’ll discuss how a dataset of abstracts and associated human-tagged keywords from a standardized list of several thousand keywords were used to create a machine-learning model that predicted keyword metadata for open-source code projects on code.nasa.gov. We’ll also discuss a less successful example from data.nasa.gov to show how data archive architecture and characteristics of initial metadata can be strong controls on how easy it is to leverage programmatic methods to reuse metadata to create additional metadata.

Gosses, Justin

Dragonfly Rotor Optimization using Machine Learning Applied to an OVERFLOW Generated Airfoil Database

NASA’s 4th New Frontiers Mission is the Titan Dragonfly relocatable lander. This coaxial quadrotor vehicle will be launched on a rocket to Titan in 2028. Following a gravity assisted Earth flyby and an approximate 6-year transit, Dragonfly will enter the Titan atmosphere around 2034 with the goal of exploring Titan’s pre-biotic chemistry and habitability. The multirotor design for this unique application has continually evolved since 2016 with constraints such as Titan’s cryogenic atmosphere at 95 Kelvin (-288 F), gravity 14% that of Earth’s, atmospheric density 440% of standard sea-level air, and the inability to test the entire system together under all these conditions until the first flight on Titan. This paper focuses on rotor design aspects of the Dragonfly lander and introduces a novel framework for multirotor design optimization considering multiple flight conditions. The methodology leverages machine learning methods and is demonstrated in the context of Dragonfly. A new OVERFLOW Machine Learning Airfoil Performance (PALMO) database is first presented. PALMO is then wrapped inside a Bayesian optimization framework and applied to a 4-rotor system (one side of the Dragonfly lander). Training data is generated on each iteration of the optimization using the CAMRAD-II comprehensive analysis software to evaluate successive rotor designs in multiple relevant flight conditions. An optimal design for the 4-rotor system was found with approximately 900 rotor designs analyzed in CAMRAD-II, which required 9 million queries of the PALMO surrogate models. This demonstration case evaluated 10,000,000 potential candidate rotor designs in 5.5 hours on 114 CPU cores using uniform inflow, and in 27.8 hours using the prescribed wake model. This work thus enables mid-fidelity rotor design optimization without requiring access to high-performance computing.

Dragonfly

Predictive Chemical Kinetic Modeling: Where We Succeed, Where We Struggle, and What Comes Next

Chemical kinetic modeling plays a foundational role in fields ranging from energy to environmental science, pharmaceuticals, and advanced materials. The past two decades have seen remarkable progress, particularly in modeling gas-phase reactions for thermochemical processes, leading to impactful industrial applications such as steam cracking and air quality management. However, new challenges are emerging. The successful development of systematic methodologies for the description of gas-phase kinetics opens the possibility to apply the same approach to the study of more challenging systems. Here, we review recent advances, including ab initio transition state theory-based master equation estimation of elementary rates, automated mechanism generation, machine-learning-assisted kinetics, and uncertainty quantification, and discuss the advances needed to apply the same methodological approach in areas such as heterogeneous catalysis, electrochemistry, liquid-phase and solid-state reactivity, and multiscale model integration. We advocate for the development of targeted tools, especially methods that go beyond empirical tuning toward first-principles-based predictions. We highlight the need for accessible software and AIaugmented workflows to democratize modeling for industry and academia alike. In this perspective, we call attention to not only what has worked but also what remains unsolved, advocating to avoid overemphasizing successes in scientific works at the expense of realism. The next decade should focus on predictive capability, physical accuracy, and community infrastructure (e.g., databases and services) to enable innovation across diverse fields. We argue that kinetic modeling, properly equipped, can accelerate discovery far beyond its traditional domains.

ab initio calculations

Reduced Erosion Augments Soil Carbon Storage Under Cover Crops

ABSTRACT Cover crops, a promising strategy to increase soil organic carbon (SOC) storage in croplands and mitigate climate change, have typically been shown to benefit soil carbon (C) storage from increased plant C inputs. However, input‐driven C benefits may be augmented by the reduction of C outputs induced by cover crops, a process that has been tested by individual studies but has not yet been synthesized. Here we quantified the impact of cover crops on organic C loss via soil erosion (SOC erosion) and revealed the geographical variability at the global scale. We analyzed the field data from 152 paired control and cover crop treatments from 57 published studies worldwide using meta‐analysis and machine learning. The meta‐analysis results showed that cover crops widely reduced SOC erosion by an average of 68% on an annual basis, while they increased SOC stock by 14% (0–15 cm). The absolute SOC erosion reduction ranged from 0 to 18.0 Mg C −1 ha −1 year −1 and showed no correlation with the SOC stock change that varied from −8.07 to 22.6 Mg C −1 ha −1 year −1 at 0–15 cm depth, indicating the latter more likely related to plant C inputs. The magnitude of SOC erosion reduction was dominantly determined by topographic slope. The global map generated by machine learning showed the relative effectiveness of SOC erosion reduction mainly occurred in temperate regions, including central Europe, central‐east China, and Southern South America. Our results highlight that cover crop‐induced erosion reduction can augment SOC stock to provide additive C benefits, especially in sloping and temperate croplands, for mitigating climate change.

Huang, Wenjuan [Department of Ecology, Evolution,

Future Building Archetypes for Los Angeles (2100 Projection)

This dataset (Data.zip) includes empirical and machine learning-generated building information for the Los Angeles urban region. The MAv1_LA.csv file provides the baseline 2015 building data while Final_IECC_LO_2100_GAN.csv represents generative adversarial network-projected urban morphologies for the year 2100. Building archetypes were created for both datasets (Basecase_LA_Archetype.csv and LA_Simulation_2100_GAN_Archetype.csv) using footprint area as the key aggregation variable. More details about the dataset are provided in the attached readme file (README_LA_Archetype_MAv1.txt)

AutoBEM

Generative models for simulation of KamLAND-Zen

Abstract The next generation of searches for neutrinoless double beta decay ($$0 \nu \beta \beta $$ 0 ν β β ) are poised to answer deep questions on the nature of neutrinos and the source of the Universe’s matter–antimatter asymmetry. They will be looking for event rates of less than one event per ton of instrumented isotope per year. To claim discovery, accurate and efficient simulations of detector events that mimic$$0 \nu \beta \beta $$ 0 ν β β is critical. Traditional Monte Carlo (MC) simulations can be supplemented by machine-learning-based generative models. This work describes the performance of generative models that we designed for monolithic liquid scintillator detectors like KamLAND to produce accurate simulation data without a predefined physics model. We present their current ability to recover low-level features and perform interpolation. In the future, the results of these generative models can be used to improve event classification and background rejection by providing high-quality abundant generated data.

Physics

Using machine learning techniques to automate sky survey catalog generation

We describe the application of machine classification techniques to the development of an automated tool for the reduction of a large scientific data set. The 2nd Palomar Observatory Sky Survey provides comprehensive photographic coverage of the northern celestial hemisphere. The photographic plates are being digitized into images containing on the order of 10(exp 7) galaxies and 10(exp 8) stars. Since the size of this data set precludes manual analysis and classification of objects, our approach is to develop a software system which integrates independently developed techniques for image processing and data classification. Image processing routines are applied to identify and measure features of sky objects. Selected features are used to determine the classification of each object. GID3* and O-BTree, two inductive learning techniques, are used to automatically learn classification decision trees from examples. We describe the techniques used, the details of our specific application, and the initial encouraging results which indicate that our approach is well-suited to the problem. The benefits of the approach are increased data reduction throughput, consistency of classification, and the automated derivation of classification rules that will form an objective, examinable basis for classifying sky objects. Furthermore, astronomers will be freed from the tedium of an intensely visual task to pursue more challenging analysis and interpretation problems given automatically cataloged data.

Fayyad, Usama M.

Utilizing Machine Learning to Improve Neutralization Potency of an HIV-1 Antibody Targeting the gp41 N-Heptad Repeat

The N-heptad repeat (NHR) of the HIV-1 gp41 prehairpin intermediate (PHI) is an attractive potential vaccine target with high sequence conservation across diverse strains. However, despite the potency of NHR-targeting peptides and clinical efficacy of the NHR-targeting entry inhibitor enfuvirtide, no potently neutralizing NHR-directed monoclonal antibodies (mAbs) nor antisera have been identified or elicited to date. The lack of potent NHR-binding mAbs both dampens enthusiasm for vaccine development efforts at this target and presents a barrier to performing passive immunization experiments with NHR-targeting antibodies. To address this challenge, we previously developed an improved variant of the NHR-directed mAb D5, called D5_AR, which is capable of neutralizing diverse tier-2 viruses. Building on that work, here we present the 2.7Å-crystal structure of D5_AR bound to NHR mimetic peptide IQN17. We then utilize protein language models and supervised machine learning to generate small (n < 100) libraries of D5_AR variants that are subsequently screened for improved neutralization potency. We identify a variant with 5-fold improved neutralization potency, D5_FI, which is the most potent NHR-directed monoclonal antibody characterized to date and exhibits broad neutralization of tier-2 and −3 pseudoviruses as well as replicating R5 and X4 challenge strains. Additionally, our work highlights the ability of protein language models to efficiently identify improved mAb variants from relatively small libraries.

Biopolymers

Insights into coordination and ligand trends of lanthanide complexes from the Cambridge Structural Database

Abstract Understanding lanthanide coordination chemistry can help develop new ligands for more efficient separation of lanthanides for critical materials needs. The Cambridge Structural Database (CSD) contains tens of thousands of single crystal structures of lanthanide complexes that can serve as a training ground for both fundamental chemical insights and future machine learning and generative artificial intelligence models. This work aims to understand the currently available structures of lanthanide complexes in CSD by analyzing the coordination shell, donor types, and ligand types, from the perspective of rare-earth element (REE) separations. We obtain four sets of lanthanide complexes from CSD: Subset 1, all Ln-containing complexes (49472 structures); Subset 2, mononuclear Ln complexes (27858 structures); Subset 3, mononuclear Ln complexes without cyclopentadienyl ligands (Cp) (26156 structures); Subset 4, Ln complexes with at least one 1,10-phenanthroline (phen) or its derivative as a coordinating ligand (2226 structures). The subsequent analysis of lanthanide complexes in these subsets examines the trends in coordination numbers and first shell distances as well as identifies and characterizes the ligands and donor groups. In addition, examples of Ln-complexes with commercially available complexants and phen-based ligands are interrogated in detail. This systematic investigation lays the groundwork for future data-driven ligand designs for REE separations based on the structural insights into the lanthanide coordination chemistry.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Conditional guided generative diffusion for particle accelerator beam diagnostics

Abstract Advanced accelerator-based light sources such as free electron lasers (FEL) accelerate highly relativistic electron beams to generate incredibly short (10s of femtoseconds) coherent flashes of light for dynamic imaging, whose brightness exceeds that of traditional synchrotron-based light sources by orders of magnitude. FEL operation requires precise control of the shape and energy of the extremely short electron bunches whose characteristics directly translate into the properties of the produced light. Control of short intense beams is difficult due to beam characteristics drifting with time and complex collective effects such as space charge and coherent synchrotron radiation. Detailed diagnostics of beam properties are therefore essential for precise beam control. Such measurements typically rely on a destructive approach based on a combination of a transverse deflecting resonant cavity followed by a dipole magnet in order to measure a beam’s 2D time vs energy longitudinal phase-space distribution. In this paper, we develop a non-invasive virtual diagnostic of an electron beam’s longitudinal phase space at megapixel resolution (1024 × 1024) based on a generative conditional diffusion model. We demonstrate the model’s generative ability on experimental data from the European X-ray FEL.

43 PARTICLE ACCELERATORS

Exploring dielectric properties in atomistic models of amorphous boron nitride

Abstract We report a theoretical study of dielectric properties of models of amorphous Boron Nitride, using interatomic potentials generated by machine learning. We first perform first-principles simulations on small (about 100 atoms in the periodic cell) sample sizes to explore the emergence of mid-gap states and its correlation with structural features. Next, by using a simplified tight-binding electronic model, we analyse the dielectric functions for complex three dimensional models (containing about 10.000 atoms) embedding varying concentrations of sp 1 , sp 2 and sp 3 bonds between B and N atoms. Within the limits of these methodologies, the resulting value of the zero-frequency dielectric constant is shown to be influenced by the population density of such mid-gap states and their localization characteristics. We observe nontrivial correlations between the structure-induced electronic fluctuations and the resulting dielectric constant values. Our findings are however just a first step in the quest of accessing fully accurate dielectric properties of as-grown amorphous BN of relevance for interconnect technologies and beyond.

Materials Science

A Method for Validating Causal Diagrams of Human Health Risk in Space Flight

The complexity of cause-and-effect relationships between spaceflight hazards and resulting health conditions clouds understanding of the totality of human system risk in space. In response, NASA has introduced Directed Acyclic Graphs (causal diagrams) into the human systems risk management process. These diagrams allow for a common understanding of the mechanisms that lead from unique hazards of spaceflight to the health outcomes important to agencies and astronauts. However, the paucity of available biomedical data from spaceflight creates a need for methods of validating causal models that can accommodate data from spaceflight model analogs. Here we outline one approach utilizing open-access rodent bone datasets from the Ames Life Sciences Data Archive. The properties of directed acyclic graphs themselves can provide an epistemological and statistical framework for validation of a priori causal representations of human system risk in space flight. The assumed causal connections on the graph creates sets of logical implications: variables that – if the causal diagram is correct – should be correlated, as well as sets that should be conditionally independent. By testing these implied correlations and conditional independencies both statistically and heuristically, we can provide evidence for or against specific causal pathways on the causal diagram. In addition to validation of expert-generated causal diagrams, machine learning techniques can learn the most likely structure of a causal diagram from a given dataset. Comparison with and reconciliation between machine-learned causal diagrams and expert-generated diagrams is another technique for challenging assumptions and improving our understanding of causal mechanisms. Accurately representing complex causation is essential to systemic understanding of human health risks in space travel. Having a robust system of validating causal diagrams helps us arrive at more accurate representations of causal systems. This process will be integral to developing the countermeasures necessary for extended exploration of the moon and Mars.

Robert Reynolds

Combining Data with Physical Knowledge for Uncertainty Quantification in Certification and Reliability Analysis

Unifying empirical data with predictive models can enable engineering cost-savings through certification by analysis and reliability-based design. Both concepts require rigorous uncertainty quantification (UQ) and robust understanding and treatment of relevant physics. Combining sampling-based UQ algorithms with high-fidelity simulations creates a computational bottleneck that is often alleviated through the use of machine learning (ML). ML can be used to create computationally efficient surrogates for simulations of complex or high-dimensional physical interactions (e.g., multi-phase interactions associated with melt pools in laser powder bed fusion or spatially-dependent material properties in functionally graded materials). However, negative side effects of ML may include a lack of interpretability and negative correlation between event rarity and simulation accuracy due to a lack of training data. As such, it is important to infuse ML algorithms with physics-based guardrails to provide confidence in their predictions. This talk will provide a brief review of recent NASA research at this intersection of physics-based simulation, ML, and UQ with a focus on certification and reliability analysis.

uncertainty quantification

Nuclear quantum effects of metal surface-mediated C–H activation

The nuclear quantum effects of surface-mediated C–H activation of surface CH 3 are considered for the pristine Pt(111) and Au(111) surfaces at 300 K. The kinetic barriers without nuclear quantum effects are calculated using both static density functional theory calculations and ab initio molecular dynamics. Static calculations are performed using the harmonic approximation while the free energy pathway is calculated using enhanced sampling molecular dynamics. Machine learning potentials are trained using generated datasets and validated against the ab initio molecular dynamics generated free energy pathways. The machine learning potentials are used to perform centroid molecular dynamics to consider the nuclear quantum effects of C–H activation. Nuclear quantum effects are found to have a very significant effect on the free energy pathway, with reduced importance at higher temperatures and in the CD 3 case.

Bunting, Rhys J. [Lawrence Livermore National Labo