Search NASA⌕ Search

SEARCH · Search NASA

Results for “workflow development”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 289 records · Page 16

Autonomous Flow Electrochemistry for Accelerated Catalyst Discovery

Our objective is to develop an Autonomous Chemical Experimentation (ACE) platform that accelerates discovery of new catalytic transformations and other energy-relevant chemical reactions and processes. We intentionally designed ACE to be highly modular, both with respect to its rapid deployment to different chemistries and experimental workflows as well as incorporation of a wide range of different AI algorithms. In addition to the development of the core software architecture, initial efforts were made to incorporate Large Language Models to provide human-interpretable reasoning of the optimizer’s actions, and to develop a user-friendly graphical interface for experimental researchers. ACE was demonstrated using a flow electrocatalysis platform containing an inline FTIR spectrometer for real-time analysis and quantification of the reaction outcome. Human-in-the-loop experiments were performed in which a human researcher conducted an experiment using electrode potentials suggested by ACE, then fed the spectral data back to ACE for decision making. After confirming the successful function of the optimizer, efforts were next directed to automation of the hardware and performed full autonomy tests using three reactions: catalytic oxidation of formate, catalytic oxidation of cyclohexanol, and oxidation of hydroquinone. These studies confirm that ACE can close the loop between reaction execution, analysis, and optimization. They also reveal that more improved product detection methods will be essential for ACE to make well-informed decisions for reactions with low conversions.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Best practices in software development for robust and reproducible geoscientific models based on insights from the Global Carbon Budget's dynamic vegetation models

Computational models play an increasingly vital role in scientific research by enabling the numerical simulation of complex processes. Such models are also fundamental in geosciences. For instance, they offer critical insights into the impacts of global change on the Earth system today and in the future. Beyond their value as research tools, models are also software products and should therefore adhere to certain established software engineering standards. However, scientists are rarely trained as software developers, which can lead to potential deficiencies in software quality like unreadable, inefficient, or erroneous code. The complexity of models, coupled with their integration into broader workflows, also often makes it challenging to reproduce results, evaluate processes, and build upon them. In this paper, we review the state and current practices of the development processes of the state-of-the-art land surface models used by the Global Carbon Budget. We combine the experience of modelers from the respective research groups with the expertise of software engineers from tech companies to outline key principles and tools for improving software quality in research. We explore four main areas: (1) model testing and validation, (2) scientific, technical, and user documentation, (3) version control, continuous integration, and code review, and (4) the portability and reproducibility of workflows. Our review reveals that while modeling communities are incorporating many best practices, significant room for improvement remains in areas such as automated testing, automated documentation, and reproducibility. Therefore, we here identify and promote essential software engineering practices, including numerous examples of practices from within the community that can serve as guidelines for other models and could help streamline processes across the entire community. We conclude with an open-source example implementation of these principles, demonstrating portable and reproducible data flows, a continuous integration setup, and web-based visualizations. This example may serve as a practical resource for model developers, users, and all scientists engaged in scientific programming.

Gregor, Konstantin [Technical Univ. of Munich (Ger↗

Unraveling the Dynamics of Nucleosome Arrays

The organization of genomic DNA into chromatin is a fundamental determinant of genome stability, regulation, and cellular function. Nucleosomes, the basic repeating units of chromatin, assemble into higher-order structures whose organization and heterogeneity remain difficult to characterize using conventional ensemble-averaged techniques. A key need in the field is the development of experimental approaches capable of directly visualizing nucleosome assemblies and their structural variability at the single-molecule level. This LDRD Lab-Wide project focused on establishing and evaluating atomic force microscopy (AFM)–based approaches for the characterization of nucleosome assemblies. The work emphasized experimental workflows for preparing, imaging, and assessing multi-nucleosome systems, rather than isolated single nucleosomes. Through method development and exploratory measurements, the project demonstrated the feasibility of applying scanning probe microscopy to investigate chromatin-relevant assemblies and provided preliminary insight into the strengths and limitations of this approach for future quantitative studies. Results and lessons learned from this effort were disseminated to the broader scientific community through multiple national conference presentations, helping to position LLNL for continued work in chromatin and genome organization research.

59 BASIC BIOLOGICAL SCIENCES↗

Redefining Design for Remanufacturing: A Practical Methodology for Prioritizing Remanufacturing Design Rules

Products are often discarded when they fail or no longer meet user needs. These outcomes are frequently shaped by early design decisions. While remanufacturing offers a sustainable alternative by restoring products to like‐new condition, its potential is often limited by designs that do not consider remanufacturing from the outset. This research addresses that challenge by introducing a structured Design for Remanufacturing (DfRem) methodology and a CAD‐integrated tool to support real‐time design decisions. The DfRem framework introduces a new primary design function focused on preserving product functionality across its life cycle. It is supported by a fault tree that identifies failure modes that limit remanufacturing potential and a hierarchy of design principles including Prevent, Minimize, Relocate, Restore, and others. Each principle is linked to actionable design rules that help engineers reduce the need for remanufacturing or improve its efficiency when necessary. To operationalize this framework, we developed CAD plugins for Autodesk Inventor and PTC Creo. These tools use a state machine model to present prioritized design rules based on selected failure modes and user input. By embedding DfRem logic directly into widely used CAD environments, the tool enables engineers to make sustainability‐informed decisions without disrupting existing workflows. Furthermore, this approach highlights the critical role of design in enabling circular and resource‐efficient product development, making remanufacturing a more practical and accessible strategy during the early stages of product design.

CAD↗

Low-Temperature Geothermal Play Fairway Analysis for the Denver Basin

This project is part of a national initiative to showcase the benefits of incorporating low-temperature geothermal resource assessment into the deployment of geothermal heating, combined heat and power (CHP), and geothermal direct-use (GDU) technologies. The initiative was established to accelerate the country's decarbonization efforts by identifying potential for low-temperature geothermal resource utilization (<150 degrees C, e.g., CHP and GDU) in selected sedimentary basins with numerous population centers. The play fairway analysis (PFA) methodologies in this study were adapted from previous PFA investigations of sedimentary basin geothermal play types (SBGPTs) that evaluated the potential for low-temperature resources (<150 degrees C). Workflows, relevant datasets, a new Python library, and common and composite geological criteria maps are utilized to develop low-temperature geothermal resource favorability maps for the Denver Basin, a sedimentary basin spanning Colorado, Nebraska, and Wyoming. The replication of these methodologies in other SBGPTs can evaluate potential for low-temperature resources. To facilitate future assessment of low-temperature geothermal resources in SBGPTs, this project provides PFA workflows, data, tools, and favorability maps that will ultimately support the utilization of low-temperature geothermal resources in sedimentary basins.

combined heat and power↗

Incorporating corrosion design constraints in desalination process optimization: A case study in mechanical vapor compression

Corrosion is an expensive and complex challenge for desalination, yet current design approaches do not explicitly account for corrosion mechanisms in process modeling and technoeconomic analysis. Here, to address this gap, we present a workflow for incorporating corrosion design constraints directly into desalination process optimization models. We develop surrogate models for general and localized corrosion metrics as functions of temperature, pH, salinity, dissolved oxygen, and material using data from OLI Systems’ Corrosion Analyzer. We then integrate these surrogates as corrosion design constraints in a cost-optimization MVC model that minimizes the levelized cost of water (LCOW). For a case study of mechanical vapor compression (MVC) treating seawater across a range of recoveries, we find dissolved oxygen (DO) is the dominant driver of localized corrosion, and thus of cost-optimal material choice and operating conditions. Reducing the DO from 8 mg/L to 0.5 mg/L reduces the LCOW by 15-35%, informing the breakeven costs for implementing DO removal or selecting highly corrosion-resistant alloys. This framework is broadly applicable across corrosion types, materials, and components and enables desalination process design that minimizes capital costs.

36 MATERIALS SCIENCE↗

The Memory Scaling of Reverse-Mode Differentiation in Particle Accelerator Simulations with Space Charge

The recent development of differentiable simulation codes for particle accelerators has enabled gradient-based workflows that promise finer control and more realistic modeling of accelerator facilities. However, when using reverse-mode automatic differentiation, the memory usage continuously increases during the simulation, and can potentially exceed the available hardware memory - especially when costly space charge computation is included. To study the memory requirements for differentiable simulations, we have implemented space charge in Cheetah, a PyTorch-based beam tracking code that supports reverse-mode differentiation. We find that the memory usage for reverse-mode differentiation grows linearly with the number of macroparticles and cells, and that it is proportional to the number of space charge kicks involved in the simulation. This general scaling can be used to evaluate whether a given differentiable simulation is feasible given hardware memory constraints.

Dhamrait, Arjun↗

Deeplynx Airflow Provider Package

The DeepLynx Airflow Provider Package is a python package used to interact with the data warehouse DeepLynx when using the workflow orchestration tool Apache Airflow. This python package is packaged together using the airflow package standard so that it can be easily installed and used in any Apache Airflow environment. This package is meant to encapsulate the DeepLynx API for use in Airflow so that any interactions with DeepLynx that a user may want to use in their Airflow workflow can be easily accomplished using this provider package. This allows us to develop, implement, and test our DeepLynx-Airflow interactions in one provider package repository, and then easily install and use this package in any airflow instance. This DeepLynx Airflow Provider Package will be used extensively by the DeepLynx DAG repository.

Cavaluzzi, JackM↗

New U.S. Data Tools are Playing a Crucial Role in Decarbonizing Buildings at Speed, Scale, and Low Cost

Preparing buildings for retrofits traditionally requires expensive on-site audits or timeintensive simulation models. As a result, the majority of buildings fail to pursue cost-saving retrofits. To address these barriers, the U.S. Department of Energy (DOE) has introduced the Building Efficiency Targeting Tool for Energy Retrofits (BETTER)-a new, free, on-line tool that utilizes a data-driven analytical engine and user-friendly web interface to automatically analyze a building's monthly energy usage in response to weather conditions. The tool benchmarks a building's electric and fossil energy usage against peers; estimates energy, cost, and emissions reductions at the building and portfolio levels; recommends energy efficiency measures; and prioritizes buildings for net-zero energy retrofits. Thanks to interoperability with the DOE's Standard Energy Efficiency Data (SEED) platform, BETTER is supporting U.S. jurisdictions to prepare buildings for retrofit at speed, scale, and low cost to comply with energy policies. This paper discusses the use of BETTER and SEED by one of the branches of the California state government to streamline a retrofit program across 455 public non-residential buildings to align with state goals to reduce greenhouse gas emissions. It describes the organization's challenge to reduce energy consumption across a geographically diverse, aging portfolio; explores how BETTER and SEED improved workflow efficiency; presents preliminary results, including avoiding audit costs of $3.28 million and developing the groundwork for retrofit projects estimated to prevent emission of 2,271 t CO2e annually; and provides guidance for other jurisdictions seeking similar results.

BETTER↗

New U.S. Data Tools are Playing a Crucial Role in Decarbonizing Buildings at Speed, Scale, and Low Cost

Preparing buildings for retrofits traditionally requires expensive on-site audits or time- intensive simulation models. As a result, the majority of buildings fail to pursue cost-saving retrofits. To address these barriers, the U.S. Department of Energy (DOE) has introduced the Building Efficiency Targeting Tool for Energy Retrofits (BETTER)—a new, free, on-line tool that utilizes a data-driven analytical engine and user-friendly web interface to automatically analyze a building’s monthly energy usage in response to weather conditions. The tool benchmarks a building’s electric and fossil energy usage against peers; estimates energy, cost, and emissions reductions at the building and portfolio levels; recommends energy efficiency measures; and prioritizes buildings for net-zero energy retrofits. Thanks to interoperability with the DOE’s Standard Energy Efficiency Data (SEED) platform, BETTER is supporting U.S. jurisdictions to prepare buildings for retrofit at speed, scale, and low cost to comply with energy policies. This paper discusses the use of BETTER and SEED by one of the branches of the California state government to streamline a retrofit program across 455 public non-residential buildings to align with state goals to reduce greenhouse gas emissions. It describes the organization’s challenge to reduce energy consumption across a geographically diverse, aging portfolio; explores how BETTER and SEED improved workflow efficiency; presents preliminary results, including avoiding audit costs of $3.28 million and developing the groundwork for retrofit projects estimated to prevent emission of 2,271 t CO2e annually; and provides guidance for other jurisdictions seeking similar results.

Li, han↗

Solar Spectrum Conversion for an Algae Bioreactor (CRADA Final Report)

This project focused on developing advanced optical coatings to improve solar energy utilization. The research aimed to create lanthanide-doped upconversion nanoparticles (UCNPs) capable of capturing unused near-infrared (NIR) light from the sun and converting it into visible light (blue and red photons) that can be used for photosynthesis. The primary goal was to identify, synthesize, and integrate highly efficient UCNPs into a transparent thin-film device. Through a comprehensive workflow involving computer simulations, high-throughput robotic synthesis, and detailed optical characterization, the project successfully developed a high-performance material. The key technical achievement was the creation of a core-shell UCNP (NaYF₄:20%Yb³⁺, 2%Er³⁺ coated with a 10 nm NaYF₄ shell) that demonstrated a quantum yield of 3.2% for converting 980 nm NIR light into visible light. Transparent thin films fabricated from these nanoparticles showed excellent optical properties, confirming their potential for practical applications. This research adds to the scientific understanding of energy transfer in lanthanide materials and demonstrates a technically effective method for creating efficient light-converting coatings. The primary benefit to the public lies in the potential for these coatings to enhance the efficiency of solar-driven processes, such as boosting the growth of algae in photobioreactors for biofuel production.

14 SOLAR ENERGY↗

Operating advanced scientific instruments with AI agents that learn on the job

Advanced scientific user facilities, such as next generation X-ray light sources and self-driving laboratories, are revolutionizing scientific discovery by automating routine tasks and enabling rapid experimentation and characterizations. However, these facilities must continuously evolve to support new experimental workflows, adapt to diverse user projects, and meet growing demands for more intricate instruments and experiments. This continuous development introduces significant operational complexity, necessitating a focus on usability, reproducibility, and intuitive human-instrument interaction. In this work, we explore the integration of agentic AI, powered by Large Language Models (LLMs), as a transformative tool to achieve this goal. We present our approach to developing a human-in-the-loop pipeline for operating advanced instruments including an X-ray nanoprobe beamline and an autonomous robotic station dedicated to the design and characterization of materials. Specifically, we evaluate the potential of various LLMs as trainable scientific assistants for orchestrating complex, multi-task workflows, which also include multimodal data, optimizing their performance through optional human input and iterative learning. We demonstrate the ability of AI agents to bridge the gap between advanced automation and user-friendly operation, paving the way for more adaptable and intelligent scientific facilities.

Large Language Models↗

Building a FAIR data ecosystem for incorporating single-cell transcriptomics data into agricultural genome to phenome research

Introduction The agriculture genomics community has numerous data submission standards available, but the standards for describing and storing single-cell (SC, e.g., scRNA- seq) data are comparatively underdeveloped. Methods To bridge this gap, we leveraged recent advancements in human genomics infrastructure, such as the integration of the Human Cell Atlas Data Portal with Terra, a secure, scalable, open-source platform for biomedical researchers to access data, run analysis tools, and collaborate. In parallel, the Single Cell Expression Atlas at EMBL-EBI offers a comprehensive data ingestion portal for high-throughput sequencing datasets, including plants, protists, and animals (including humans). Developing data tools connecting these resources would offer significant advantages to the agricultural genomics community. The FAANG data portal at EMBL-EBI emphasizes delivering rich metadata and highly accurate and reliable annotation of farmed animals but is not computationally linked to either of these resources. Results Herein, we describe a pilot-scale project that determines whether the current FAANG metadata standards for livestock can be used to ingest scRNA-seq datasets into Terra in a manner consistent with HCA Data Portal standards. Importantly, rich scRNA-seq metadata can now be brokered through the FAANG data portal using a semi-automated process, thereby avoiding the need for substantial expert curation. We have further extended the functionality of this tool so that validated and ingested SC files within the HCA Data Portal are transferred to Terra for further analysis. In addition, we verified data ingestion into Terra, hosted on Azure, and demonstrated the use of a workflow to analyze the first ingested porcine scRNA-seq dataset. Additionally, we have also developed prototype tools to visualize the output of scRNA-seq analyses on genome browsers to compare gene expression patterns across tissues and cell populations. This JBrowse tool now features distinct tracks, showcasing PBMC scRNA-seq alongside two bulk RNA-seq experiments. Discussion We intend to further build upon these existing tools to construct a scientist-friendly data resource and analytical ecosystem based on Findable, Accessible, Interoperable, and Reusable (FAIR) SC principles to facilitate SC-level genomic analysis through data ingestion, storage, retrieval, re-use, visualization, and comparative annotation across agricultural species.

Genetics & Heredity↗

Towards an Introspective Dynamic Model of Globally Distributed Computing Infrastructures

Large-scale scientific collaborations like ATLAS, Belle II, CMS, DUNE, and others involve hundreds of research institutes and thousands of researchers spread across the globe. These experiments generate petabytes of data, with volumes soon expected to reach exabytes. Consequently, there is a growing need for computation, including structured data processing from raw data to consumer-ready derived data, extensive Monte Carlo simulation campaigns, and a wide range of end-user analysis. To manage these computational and storage demands, centralized workflow and data management systems are implemented. However, decisions regarding data placement and payload allocation are often made disjointly and via heuristic means. A significant obstacle in adopting more effective heuristic or AI-driven solutions is the absence of a quick and reliable introspective dynamic model to evaluate and refine alternative approaches. In this study, we aim to develop such an interactive system using real-world data. By examining job execution records from the PanDA workflow management system, we have pinpointed key performance indicators such as queuing time, error rate, and the extent of remote data access. The dataset includes five months of activity. Additionally, we are creating a generative AI model to simulate time series of payloads, which incorporate visible features like category, event count, and submitting group, as well as hidden features like the total computational load—derived from existing PanDA records and computing site capabilities. These hidden features, which are not visible to job allocators, whether heuristic or AI-driven, influence factors such as queuing times and data movement.

kilic, Ozgur Ozan [Brookhaven National Laboratory ↗

UrbanScaping: Community Spatial Data Visualization & Analytics

Evaluating the electrification potential of buildings through retrofitting is crucial for reducing carbon emissions and the carbon footprint of built environments. This study leverages the Automatic Building Energy Modeling (AutoBEM) software, integrating the Model America database to create an urban context-based spatial analysis platform for community engagement and development. We selected Camp Hill Borough, PA, as a case study to analyze building-specific energy performance and evaluate the electrification potential of each building by switching to different Heating, Ventilation, and Air Conditioning (HVAC) systems and measurement components. The simulation results generated by the workflow provide retrofitting suggestions to help mitigate the carbon footprint as well as energy saving statistics of buildings. Additionally, the developed web-based interface serves as a community engagement platform, allowing residents to provide feedback and further develop interactive communication protocols. The outcomes of this project offer a baseline for community electrification planning and contribute to the design of low-carbon communities.

Chowdhury, Shovan [ORNL]↗

CI-MOR Final Report: Analysis and Validation of Critical Infrastructure Models using Model Order Reduction

This report summarizes the research and capabilities developed as part of the project “Analysis and Validation of Critical Infrastructure Models using Model Order Reduction” (CI-MOR) LDRD project. CI-MOR research enables the solution of large, complex optimization models that naturally arise in national security challenges involving critical infrastructures. Specifically, CI-MOR researchers developed methods to (1) rigorously approximate complex, nonlinear optimization formulations, (2) identify alternative near-optimal solutions, (3) accelerate optimization workflows used for complex applications, and (4) rigorously integrate domain knowledge in stochastic-process models. This report provides an overview of the research done in CI-MOR, and we describe application exemplars used to illustrate CI-MOR capabilities. Furthermore, we describe the software developed by CI-MOR that researchers can leverage to analyze new applications.

97 MATHEMATICS AND COMPUTING↗

Tracking Dendritic Growth in Hydrogen-Based Hematite Reduction via Computer Vision

The reduction of hematite to metallic iron using hydrogen (H2) as a reducing agent presents a promising pathway for decarbonizing steel production. In this study, we employ a combination of in situ confocal scanning laser microscopy (CSLM) and advanced computer vision techniques to quantitatively analyze dendritic growth of ferrite during H2-based reduction of iron oxide at high temperatures. A workflow integrating Watershed Image Segmentation (WIS) and Lucas-Kanade Optical Flow (LKOF) is developed to extract both global and local kinetic information from time-resolved micrograph sequences. H2 reduction experiments conducted at 1400 degrees C and 1500 degrees C demonstrate a clear correlation between temperature and reduction rate, as evidenced by accuracy of fitted Johnson-Mehl-Avrami-Kolmogorov (JMAK) parameters. Optical flow analysis further elucidates the anisotropic and branched nature of dendritic growth, providing spatially resolved velocity fields that correlate well with global transformation kinetics. The proposed methodology demonstrates strong agreement with experimental measurements and literature values, offering a robust framework for automated image-based analysis to study kinetics through microstructural evolution in the reduction of iron ore, and likely other reaction-diffusion phenomena.

08 HYDROGEN↗

Assessing the numerical stability of physics models to equilibrium variation through database comparisons on DIII-D

High fidelity kinetic equilibria are crucial for tokamak modeling and analysis. Manual workflows for constructing kinetic equilibria are time consuming and subject to user error, motivating development of automated equilibrium reconstruction tools to provide accurate and consistent reconstructions for downstream physics analysis. These automated tools also provide access to kinetic equilibria at large database scales, which enables the quantification of general uncertainties arising from equilibrium reconstruction techniques. In this paper, we compare a large database of DIII-D kinetic equilibria generated manually by physics experts to equilibria from automated kinetic reconstruction tools, assessing the impact of reconstruction method on equilibrium parameters and resulting magnetohydrodynamic stability calculations. We find agreement among scalar parameters, whereas profile quantities, such as the bootstrap current, show larger disagreements. We analyze ideal kink and classical tearing stability with DCON and STRIDE respectively, finding that the kink stability calculation is generally more robust than the tearing index Δ' calculation. We find that in 90% of cases, both kink stability classifications are unchanged between the manual expert and automated kinetic equilibria.

CAKE↗