Search NASA⌕ Search

SEARCH · Search NASA

Results for “AI for Science”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 361 records · Page 20

Water4Energy Tier-1 Raw Observations for TVA Seasonal Prediction, Version 1

Tier-1 (Step-1) raw observation staging collection for the Water4Energy Genesis Task-1 project on weeks-to-years prediction of Tennessee Valley temperature and precipitation. This data-only deposit includes CPC/PSL teleconnection indices, NOAA OISST monthly and ERSST sea-surface temperature, ERA5-derived daily 1° fields (t2m, tp, msl for 1980–2024), CFSv2 NMME ensemble-mean seasonal baselines, and a TVA boundary mask. The product supports seasonal teleconnection diagnostics and construction of AI ready-to-train packs published separately. Multi-terabyte hourly archives are excluded from this version.

54 ENVIRONMENTAL SCIENCES↗

Reticular Materials and AI-Driven Computer Simulations for Seawater Mining of Valuable Metals (Final Technical Report)

This Final Technical Report describes our exploratory efforts that combine reticular materials synthesis (hydrolytically robust metal–organic frameworks, MOFs) with AI‑enabled molecular simulations to develop mechanistic, quantitative design rules for recovering lithium and other alkali-metal ions from highly dilute, competitive aqueous resources (e.g., seawater). The central outcome is a joint experimental–computational study of ion uptake in MOF‑808 (Chemical Science, 2025) that quantifies both thermodynamics and kinetics of Li + , Na + , and K + uptake and identifies how pore size, pore hydration state, dehydration penalties, and pore-window transport barriers govern selectivity. Guided by these insights, we synthesized and tested functionalized MOF‑808 and multivariate MOFs incorporating ion-recognition motifs (including carboxylates and crown-ether linkers) and evaluated uptake in synthetic seawater, highlighting framework topology and pore chemistry as levers for improved Li + /Na + discrimination. We also developed transferable simulation models, enhanced-sampling protocols, and automated workflows that enable systematic screening of porous sorbents.

42 ENGINEERING↗

Satellite and Reanalysis Air Quality Data and Services at NASA GES DISC for Public Health Study

Outbreaks of infectious diseases and health can be influenced by airborne and water-borne pollutants. Furthermore, air and water quality are associated with climate variability, industrialization, land use and land cover change, and water resource management. It is therefore crucial to understand environment-disease connections with existing long-term observed and modeled data, particularly for development of early warning systems for infectious disease outbreaks. The NASA Goddard Earth Sciences Data and Information Services Center (GES DISC) (https://disc.gsfc.nasa.gov) archives large volumes of global environment data that are useful for research and applications regarding environmental factors and public health. Examples of air quality measurements are: Daily satellite remotely sensed data, including Aerosol Index (AI), O3, SO2, CO, and NO2 from Aura/OMI (October 2014 to present), and OMPS-NPP (January 2012 to present, currently research data products only) Hourly and monthly reanalysis modeled data, including PM2.5, O3, CO, SO2, BC, dust, AOD, and aerosol types from MERRA-2 (January 1980 to present) Examples of surface meteorology and land surface measurements are: hourly, daily, and monthly satellite precipitation from TRMM (December 1997 to March 2015) 30-minute and monthly satellite precipitation from GPM (March 2014 to present); Hourly and monthly modeled surface meteorology and land surface condition from MERRA-2 (January 1980 to present) and land surface assimilation models (January 1948 to present), including precipitation, surface temperature, relative humidity, wind, and soil moisture.This presentation will give an overview of relevant environmental data at the NASA GES DISC. Through a number of use cases, such as dust events and active fires, we will introduce data services that assist in finding the right data, enable visualization and analysis of the data online, and allow downloading of data in user-preferred format.

remote sensing data↗

Environmentally Assisted Fatigue in Light Water Reactor Environment

This report summarizes the Environmentally Assisted Fatigue (EAF) research conducted at ANL under the US DOE Light Water Reactor Sustainability (LWRS) program. Starting from a rich background in theoretical and experimental EAF, ANL previously developed an approach to evaluate fatigue performance of reactor materials in light water reactor environments with the correction factor F en . The approach was based on a large body of experimental work performed at ANL and elsewhere, and was consistent with American Society of Mechanical Engineers (ASME)’s methodology governing the design and construction of reactor components. In recent years, the program was focused on component fatigue prediction and made several major and fundamental contributions in this area. These accomplishments help meet the needs identified by the industry concerning component level fatigue predictions in complex, transient conditions. The main contribution of the ANL program involved the development of a system-level model for estimating residual strain and life of nuclear reactor coolant system components under connected-system-thermal-mechanical boundary conditions. The goal was to predict the stress hotspots, strain residuals, strain amplitudes and the resulting fatigue lives. Thermal-mechanical stress analysis was performed considering thermal stratification and a design-basis reactor loading cycle. Based on the finite element (FE) model results, the strain residuals, strain amplitudes and resulting fatigue lives of reactor coolant system (RCS) components were predicted. The results show that some of the RCS components can have significantly different strain amplitudes, residual strain, and fatigue lives, despite having similar geometry and material. In addition, the simulated component-level strain profile can guide the selection of appropriate test inputs for conducting laboratory-scale EAF tests. Building upon the system-level model, ANL developed a digital twin (DT) framework to predict the structural states and associated fatigue life of components in real-time. This framework is a comprehensive system designed to predict the structural states and fatigue lives of reactor components. It includes multiple models and integrates artificial intelligence (AI), machine learning (ML), and FE based modeling tools to evaluate the structural states and fatigue lives.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Measurement of the Born cross section for 𝑒 +⁢ 𝑒 − → 𝜂⁢ℎ 𝑐 at center-of-mass energies between 4.1 and 4.6 GeV

We measure the Born cross section for the reaction 𝑒 + ⁢𝑒 − → 𝜂⁢ℎ 𝑐 from $\sqrt{𝑠}$ = 4.129 to 4.600 GeV using datasets collected by the BESIII detector running at the BEPCII collider. A resonant structure in the cross-section line shape near 4.200 GeV is observed with a statistical significance of 7⁢𝜎. The parameters of this resonance are measured to be 𝑀 = 4188.8 ± 4.7 ± 8.0 MeV/𝑐 2 and Γ = 49 ± 16 ± 19 MeV, where the first uncertainties are statistical and the second systematic.

exotic mesons↗

Artificial intelligence in a mission operations and satellite test environment

A Generic Mission Operations System using Expert System technology to demonstrate the potential of Artificial Intelligence (AI) automated monitor and control functions in a Mission Operations and Satellite Test environment will be developed at the National Aeronautics and Space Administration (NASA) Jet Propulsion Laboratory (JPL). Expert system techniques in a real time operation environment are being studied and applied to science and engineering data processing. Advanced decommutation schemes and intelligent display technology will be examined to develop imaginative improvements in rapid interpretation and distribution of information. The Generic Payload Operations Control Center (GPOCC) will demonstrate improved data handling accuracy, flexibility, and responsiveness in a complex mission environment. The ultimate goal is to automate repetitious mission operations, instrument, and satellite test functions by the applications of expert system technology and artificial intelligence resources and to enhance the level of man-machine sophistication.

Busse, Carl↗

AI Driven Optimization of Public Transit

This project explores the application of AI-driven methods to optimize public transit operations for the Chattanooga Area Regional Transportation Authority (CARTA). By leveraging data analytics, machine learning, and predictive modeling, the initiative seeks to enhance system efficiency, improve rider experience, and support sustainability goals. This research, supported by the National Science Foundation and the U.S. Department of Energy, integrates real-time transit data with advanced computational tools to inform decision-making, optimize routes, and balance operational demands. The work exemplifies a forward-looking model for mid-sized cities aiming to modernize mobility systems through intelligent technology integration.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Grid Enabled Geospatial Catalogue Web Service

Geospatial Catalogue Web Service is a vital service for sharing and interoperating volumes of distributed heterogeneous geospatial resources, such as data, services, applications, and their replicas over the web. Based on the Grid technology and the Open Geospatial Consortium (0GC) s Catalogue Service - Web Information Model, this paper proposes a new information model for Geospatial Catalogue Web Service, named as GCWS which can securely provides Grid-based publishing, managing and querying geospatial data and services, and the transparent access to the replica data and related services under the Grid environment. This information model integrates the information model of the Grid Replica Location Service (RLS)/Monitoring & Discovery Service (MDS) with the information model of OGC Catalogue Service (CSW), and refers to the geospatial data metadata standards from IS0 19115, FGDC and NASA EOS Core System and service metadata standards from IS0 191 19 to extend itself for expressing geospatial resources. Using GCWS, any valid geospatial user, who belongs to an authorized Virtual Organization (VO), can securely publish and manage geospatial resources, especially query on-demand data in the virtual community and get back it through the data-related services which provide functions such as subsetting, reformatting, reprojection etc. This work facilitates the geospatial resources sharing and interoperating under the Grid environment, and implements geospatial resources Grid enabled and Grid technologies geospatial enabled. It 2!so makes researcher to focus on science, 2nd not cn issues with computing ability, data locztic~, processir,g and management. GCWS also is a key component for workflow-based virtual geospatial data producing.

Chen, Ai-Jun↗

Keeping LAMMPS cutting edge

Since its inception 30 years ago, LAMMPS has grown to be a world-class molecular dynamics code and a cornerstone of computational materials science research. This project aimed to keep LAMMPS at the forefront of molecular dynamics simulations by adapting LAMMPS to the latest developments in machine learning technology and hardware. Initially, the project set out to provide a unified implementation of active learning for efficient training data generation in LAMMPS, but the research trajectory pivoted to address more immediate and impactful opportunities. On the hardware side, recent record-breaking molecular dynamics simulations were developed on the Cerebras wafer-scale AI chip, and this project has developed an interface between LAMMPS and the hardware-specific molecular dynamics code to accelerate and simplify development and user adoption. On the software side, PyTorch’s Ahead-of-Time (AOT) compilation features promised increased performance for state-of-the-art equivariant neural network potentials, and this project laid the groundwork for their adoption in LAMMPS, resulting in a nearly 20x acceleration in extreme cases. Combined with a comprehensive benchmark study of LAMMPS across all current exascale systems, this project has reinforced LAMMPS’s role as a versatile, high-performance tool for current and future materials science applications.

36 MATERIALS SCIENCE↗

Genesis Data Card Schema, Template and Supporting Tools

Genesis Data Cards provide a standardized template and schema for documenting scientific datasets in support of discovery, access, interoperability, reusability, governed use, and AI usability. This release of the Genesis Data Card repository includes a versioned Markdown template, a LinkML schema with generated Pydantic and JSON artifacts, schema documentation, and example completed data cards. Validation tooling is provided to ensure that completed data cards conform to the schema prior to submission. Accompanying documentation for the structured metadata is provided as a Field Reference Guide. The schema and accompanying template provided in this repository address the call for actionable context that enables humans and AI systems to find, access, interpret, cite, and reuse data, and, when appropriate, integrate it into AI and machine learning workflows. The data card is intended to serve as a common metadata artifact intended to support standardized, cross-program dataset documentation across Department of Energy (DOE)-aligned efforts, including but not limited to Genesis Mission-related implementations, the Office of Science, National Nuclear Security Administration (NNSA), and Advanced Simulation and Computing (ASC) data governance and stewardship initiatives.

data card↗

Neutrons in Structural Biology: Challenges and Opportunities (Workshop Report)

Gaining a thorough understanding of biological systems requires building our knowledge about biological processes from the level of atoms and electrons, and up to whole organisms. Such comprehensive knowledge will allow for a predictive understanding of complex biological systems behavior. It will guide us in the design and development of novel therapeutics and vaccines to tackle existing health threats and to prepare for future pandemics, and it will provide information necessary to create new biomaterials and bio-inspired technologies through manipulation of biological macromolecules, their assemblies, single cells and even microorganisms. Reaching these goals will require a synergistic combination of multiple experimental techniques with molecular calculations and predictive simulations, and the design and development of new techniques and capabilities that bridge current knowledge and technology gaps. Neutron scattering provides unique information about the biomacromolecular structure and function and can play a major role in achieving these goals. A workshop was held to engage the scientific community in identifying pressing challenges in biochemistry, structural biology, enzymology and structure-guided drug design not solved with the current neutron scattering technologies or utilizing other structural biology techniques such as X-ray crystallography, NMR, and cryo-EM. The workshop brought together structural biology, biochemistry and computational experts, as well as early career researchers and students, creating a forum for discussing scientific advancement and collaboration. The workshop included a one-day satellite training workshop where graduate students and postdoctoral researchers were educated in the application of neutron crystallography and small-angle scattering in structural biology. Furthermore, the Instrument Scientific Advisory Board (ISAB) for the development of a macromolecular neutron diffractometer at ORNL’s Second Target Station was introduced at the workshop. The major outcome was that neutrons can provide atomic-level understanding of biomacromolecular structure, function and dynamics which is of paramount importance for addressing the identified challenges. Neutron crystallography, in particular, can resolve long-standing biochemical issues regarding enzyme function by delineating the underlying chemistry and can have a major impact on the design of small-molecule therapeutics, especially in combination with molecular computation (quantum chemistry and molecular dynamics simulations) and the emerging artificial intelligence (AI)-assisted drug design technologies. The unique properties of neutrons, including their high sensitivity to hydrogen and their non-destructive nature, make them ideal probes of biological matter. There is a palpable need in the scientific community to expand and enhance the impact of neutron sciences on biology. Neutron crystallography is the only structural biology method capable of determining positions of all hydrogen atoms in proteins, nucleic acids and their complexes at near-physiological temperatures and of unstable species at cryogenic temperatures. Moreover, neutron analysis is non-ionizing, non-destructive and does not perturb the structure or redox chemistry of active site metal centers and clusters in proteins, which can be invaluable for studying radiation-sensitive metalloprotein complexes. Further, neutron energies used in scattering applications are similar to atomic motions, permitting neutron spectroscopies to characterize the dynamics of biomacromolecules on the picosecond to microsecond timescales. The different sensitivities of neutrons to protium (H) and deuterium (D) isotopes of hydrogen allow enhanced visibility of specific parts of biological complexes through isotopic labeling. The impact of neutrons will be most powerful when neutron scattering is combined with complementary experimental techniques that use photons and electrons, and with high-performance computing. The interconnection and mutuality of the experimental and theoretical capabilities will drive discoveries in biological and health sciences to generate more complete picture of complex biological systems. The major limitation in the field of biological neutron crystallography has been signal-to-noise, demanding large samples that are difficult to produce for the majority of biomacromolecules and limiting the applicability of this technique in biological sciences. A neutron crystallography instrument at the Second Target Station will revolutionize biological science with neutrons by engaging a large scientific community of structural biologists, enabling successful neutron diffraction experiments from radically smaller biomacromolecular crystals, resolving unanswered biochemical questions, and meaningfully contributing to rational drug design. The meeting highlighted 10 grand challenges that will be addressed with this advanced capability over the next decade and beyond, and the recommendations required to help address them are given below.

59 BASIC BIOLOGICAL SCIENCES↗

FAIR Framework for Physics-Inspired AI in High Energy Physics (Final Technical Report)

The main deliverable of this proposal was to publish data from high energy physics experiments in a FAIR format so that non-specialists could develop machine learning technologies using our data. The Minnesota team of Profs. Cushman, Furmanski and Rusack, from the high energy experiments CDMS, Micro-Boone and CMS, respectively, and Prof J. Sun from Computer Science worked to organize the data, to provide code to access the data, and where relevant provide documentation describing the data. The FAIR4HEP collaboration was formed with groups from UC San Diego, MIT, and the University of Illinois, with the principal investigator was Dr. Huerta. Collectively we collaborated on the publication of datasets from the LHC experiments. Members of the Minnesota group contributed to the common papers published by the collaboration

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Beyond sequence similarity: toward function-based screening of nucleic acid synthesis

Synthetic nucleic acids are a key input to modern biotechnology, yet they represent dual-use materials that require robust screening to mitigate biosecurity risks. The prevailing screening paradigm, which identifies sequences of concern (SoCs) through sequence similarity to controlled pathogens and toxins, may not fully capture risks posed by AI tools that can decouple biomolecular function from reliance on known sequences. Rapidly advancing biodesign capabilities enable the generation of genes and proteins that might evade sequence-based detection. We highlight the critical need for function-based screening approaches that can detect sequences capable of hazardous biological functions, regardless of similarity to known SoCs. We examine the feasibility of function-based screening with an initial focus on proteins, arguing that, while protein sequence space is vast, biologically functional proteins are significantly constrained by biophysical and biochemical requirements that can be learned and modeled. We propose a concrete implementation framework organized along a continuum of complexity, starting with toxins as the most tractable targets before expanding to more complex pathogenic functions. We then discuss open challenges and describe a research and development strategy to address them.

59 BASIC BIOLOGICAL SCIENCES↗

Evaluating the Effectiveness of Retrieval-Augmented Large Language Models in Scientific Document Reasoning

Despite the dramatic progress in Large Language Model (LLM) development, LLMs often provide seemingly plausible but not factual information, often referred as hallucinations. Retrieval-augmented LLMs provide a non-parametric approach to solve these issues by retrieving relevant information from external data sources and augment the training process. These models helps to trace evidence from an externally provided knowledge base allowing the model predictions to be better interpreted and verified. In this work, we critically evaluate these models in their ability to perform in scientific document reasoning tasks. To this end, we tuned multiple such model variants with science-focused instructions and evaluated them on a scientific document reasoning benchmark for the usefulness of the retrieved document passages. Our findings suggest that models justify predictions in science tasks with fabricated evidence and leveraging scientific corpus as pretraining data does not alleviate the risk of evidence fabrication.

• Artificial intelligence (AI) / machine learning ↗