Search NASA⌕ Search

SEARCH · Search NASA

Results for “DAG”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Deeplynx Dag Repository

The DeepLynx DAG repository will contain several Airflow DAGs (Directed Acyclic Graphs) which will be used in the context of DeepLynx's deployed Apache Airflow instance. These DAGs will be used for multiple data management tasks for DeepLynx data, including but not limited to: - bringing data from various sources and tools into DeepLynx - managing sequential data workflows, such as running Python scripts on data to perform analysis and returning the results to DeepLynx - performing any necessary transformation or pre-processing on data coming into DeepLynx from external sources or out of DeepLynx to go to external applications

Brownlee, JarenM.↗

Diacylglycerol enantiomer selectivity of diacylglycerol acyltransferases highlights metabolic specialization in triacylglycerol synthesis across the tree of life

Triacylglycerols are the major energy storage lipids in plants, animals, and microorganisms, and are predominantly produced by acyl-CoA:diacylglycerol (DAG) acyltransferases (DGATs). Two enantiomers of the DAG substrate, sn -1,2 and sn -2,3, can be produced by different biological mechanisms; however, little is known about which species produce each enantiomer, the selectivity of DGAT isoforms for either enantiomer, or whether DGAT enantiomer selectivity varies across organisms. Here, DAG enantiomer selectivity of DGAT1 and DGAT2 was measured from eight seed plants, two mammals, one oleaginous yeast, and one photosynthetic microalga using enantiomer-specific in vitro DGAT assays. Across most plants, DGAT1 favored sn -1,2-DAG, whereas DGAT2 preferentially utilized sn -2,3-DAG. However, there were several exceptions. Mammalian DGAT1, DGAT2, and microbial DGAT1s efficiently used both DAG enantiomers, while microbial DGAT2s had unique selectivity. The selectivity of several DGATs for combined acyl-CoA and DAG enantiomer molecular species were also evaluated for biotechnical applications. Therefore, DGAT DAG enantiomer selectivity is common yet strongly dependent on lineage and isoform and likely shaped in part by species-specific metabolic context of triacylglycerol synthesis, turnover, and remodeling. This work expands our understanding of DGAT function and establishes a foundation for leveraging enantiomer-selective acyltransferases in metabolic engineering of tailored lipid products.

Arabidopsis thaliana↗

Targeted engineering of camelina and pennycress seeds for ultrahigh accumulation of acetyl-TAG

Acetyl-TAG (3-acetyl-1,2-diacylglycerol), unique triacylglycerols (TAG) possessing an acetate group at the sn -3 position, exhibit valuable properties, such as reduced viscosity and freezing points. Previous attempts to engineer acetyl-TAG production in oilseed crops did not achieve the high levels found in naturally producing Euonymus seeds. Here, we demonstrate the successful generation of camelina and pennycress transgenic lines accumulating nearly pure acetyl-TAG at 93 mol% and 98 mol%, respectively. These ultrahigh acetyl-TAG synthesizing lines were created using gene-edited FATTY ACID ELONGASE1 ( FAE1 ) mutant lines as an improved genetic background to increase levels of acetyl-CoA available for acetyl-TAG synthesis mediated by the expression of EfDAcT, a high-activity diacylglycerol acetyltransferase isolated from Euonymus fortunei . Combining EfDAcT expression with suppression of the competing TAG-synthesizing enzyme DGAT1 further enhanced acetyl-TAG accumulation. These ultrahigh levels of acetyl-TAG exceed those in earlier engineered oilseeds and are equivalent or greater than those in Euonymus seeds. Imaging of lipid localization in transgenic seeds revealed that the low amounts of residual TAG were mostly confined to the embryonic axis. Similar spatial distributions of specific TAG and acetyl-TAG molecular species, as well as their probable diacylglycerol (DAG) precursors, provide additional evidence that acetyl-TAG and TAG are both synthesized from the same tissue-specific DAG pools. Remarkably, this ultrahigh production of acetyl-TAG in transgenic seeds exhibited minimal negative effects on seed properties, highlighting the potential for production of designer oils required for economical biofuel industries.

09 BIOMASS FUELS↗

Understanding the dynamic nature of plant lipid anabolic and catabolic metabolism is key to sustainable oilseed engineering

Plant-derived oils are essential sources of reduced carbon and various fatty acid (FA) structures for food, biofuels, and the oleochemical industry. Despite extensive efforts, engineering mainstream oilseed crops to produce high levels of industrially valuable unusual FAs (UFAs) remains challenging. This review synthesizes recent advances in the understanding of lipid metabolic networks, emphasizing how species-specific regulation of FA synthesis, activation, and delivery influences triacylglycerol (TAG) assembly to govern the efficiency of UFA accumulation. Key insights reveal that acyl flux through anabolic and catabolic branches of lipid metabolism is tightly controlled by enzyme substrate selectivities, diacylglycerol (DAG) pool compartmentalization, and metabolic context, including lipid remodeling and degradation pathways. Engineering success is often constrained by incompatibilities between UFA biosynthetic enzymes and endogenous host metabolism, leading to flux imbalances, futile cycles, and undesired phenotypes. We highlight emerging strategies to overcome these barriers, such as the use of UFA-selective acyltransferases, coordinated manipulation of DAG source pools, suppression of competing endogenous enzymes, and exploitation of TAG remodeling mechanisms. This integrated synthesis provides a conceptual framework for logic-based engineering of oilseeds with enhanced UFA content by offering new avenues for sustainable biomanufacturing of valuable lipids.

acyltransferase specificity↗

Metabolomic and transcriptomic remodeling of bone marrow myeloid cells in response to maternal obesity

Maternal obesity puts the offspring at high risk of developing obesity and cardiometabolic diseases in adulthood. Here, we utilized a mouse model of maternal high-fat diet (HFD)-induced obesity that recapitulates metabolic perturbations seen in humans. We show increased adiposity in the offspring of HFD-fed mothers (Off-HFD) when compared with the offspring of regular diet-fed mothers (Off-RD). We have previously reported significant immune perturbations in the bone marrow of newly weaned Off-HFD. Here, we hypothesized that lipid metabolism is altered in the bone marrow of Off-HFD versus Off-RD. To test this hypothesis, we investigated the lipidomic profile of bone marrow cells collected from 3-week-old Off-RD and Off-HFD. Diacylglycerols (DAGs), triacylglycerols (TAGs), sphingolipids, and phospholipids were remarkably different between the groups, independent of fetal sex. Levels of cholesteryl esters were significantly decreased in Off-HFD, suggesting reduced delivery of cholesterol. These were accompanied by age-dependent progression of mitochondrial dysfunction in bone marrow cells. We subsequently isolated CD11b+ myeloid cells from 3-wk-old mice and conducted metabolomic, lipidomic, and transcriptomic analyses. The lipidomic profiles of myeloid cells were similar to those of bone marrow cells and included increases in DAGs and decreased TAGs. Transcriptomics revealed altered expression of genes related to immune pathways, including macrophage alternative activation, B-cell receptors, and transforming growth factor-β signaling. All told, this study revealed lipidomic, metabolomic, and gene expression abnormalities in bone marrow cells broadly, and in bone marrow myeloid cells particularly, in the newly weaned offspring of mothers with obesity, which might at least partially explain the progression of metabolic and cardiovascular diseases in their adulthood.

RNA sequencing↗

OpenSn: A massively parallel, open-source simulation environment for discrete ordinates radiation transport

OpenSn is an open-source, massively parallel deterministic radiation transport code for solving the discrete-ordinates ( S N ) form of the Boltzmann transport equation on unstructured, arbitrary polyhedral meshes. It supports high-fidelity simulations involving steady-state, eigenvalue, and adjoint problems for neutral particles (e.g., neutrons, photons, multi-particles), using the multigroup approximation in energy. OpenSn combines angular discretization via discrete ordinates with a discontinuous Galerkin finite element method (DGFEM) in space, enabling accurate resolution of transport physics on arbitrary polyhedral cells, included locally refined spatial grids. It includes multiple angular quadrature types, including locally refined angular quadratures. Written in modern C++ with a Python API, OpenSn runs efficiently on platforms ranging from laptops to supercomputers. The transport sweep algorithm is implemented using a task-based, directed-acyclic-graph (DAG) approach for each angle and supports asynchronous parallelism across thousands of MPI ranks. Group-set aggregation improves compute intensity, and synthetic acceleration techniques (e.g., diffusion synthetic acceleration, second-moment method) enhance solver convergence. OpenSn has been verified on reactor physics problems and demonstrated excellent weak and strong scaling performance on more than 32,768 processes, making it a versatile and robust platform for large-scale transport simulations in complex geometries.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Integrating the PanDA Workload Management System with the Vera C. Rubin Observatory

The Vera C. Rubin Observatory will produce an unprecedented astronomical data set for studies of the deep and dynamic universe. Its Legacy Survey of Space and Time (LSST) will image the entire southern sky every three to four days and produce tens of petabytes of raw image data and associated calibration data over the course of the experiment’s run. More than 20 terabytes of data must be stored every night, and annual campaigns to reprocess the entire dataset since the beginning of the survey will be conducted over ten years. The Production and Distributed Analysis (PanDA) system was evaluated by the Rubin Observatory Data Management team and selected to serve the Observatory’s needs due to its demonstrated scalability and flexibility over the years, for its Directed Acyclic Graph (DAG) support, its support for multi-site processing, and its highly scalable complex workflows via the intelligent Data Delivery Service (iDDS). PanDA is also being evaluated for prompt processing where data must be processed within 60 seconds after image capture. This paper will briefly describe the Rubin Data Management system and its Data Facilities (DFs). Finally, it will describe in depth the work performed in order to integrate the PanDA system with the Rubin Observatory to be able to run the Rubin Science Pipelines using PanDA.

79 ASTRONOMY AND ASTROPHYSICS↗

Preparation of the Multi-Site Data Processing at the Vera C. Rubin Observatory

The Vera C. Rubin Observatory’s Legacy Survey of Space and Time (LSST) Camera is scheduled to start taking data in the summer of 2025. The Data Release Production will run the LSST Science Pipe software at data facilities in the US, France and the UK. The LSST Science Pipeline consists of complex directed acyclic graphs (DAGs) of tasks. Rubin will use the Production and Distributed Analysis (PanDA) workflow and workload management system to orchestrate this complex workflow and the distribution of workloads to the data facilities. When run end-to-end by a team of data production staff, this processing (the Science Pipelines, distributed by the workflow and workload management system) is referred to as a 'campaign'. This paper describes the central services and data facility specific services that support this multi-site data process model, including the service deployment infrastructure, the workload and workflow system, the Campaign Management tools, and connection to Rubin Data Management. This paper will also mention the experience of processing the Rubin Commissioning Camera data. All these are part of the effort to scale up the processing capabilities for the expected very large data volume from the LSST Camera.

Yang, Wei [SLAC]↗

Arabidopsis lipins mediate lipid droplet biogenesis to protect cells from lipotoxicity

Lipin proteins, a family of phosphatidic acid phosphatases (PAHs), are key regulators of lipid metabolism, storage, and homeostasis across eukaryotes. While Arabidopsis (Arabidopsis thaliana) lipins function in lipid biosynthesis and gene regulation, their roles in lipid droplet (LD) biogenesis and lipid homeostasis remain largely unknown. Here, we show that double knockout of two PAH genes (PAH1/2) results in impaired LD biogenesis, accelerated triacylglycerol (TAG) hydrolysis, and lipid imbalance. pah1/2 mutant leaves exhibited a marked reduction in TAG levels and a significant decrease in LD size, while the rates of TAG and diacylglycerol (DAG) synthesis remained largely unchanged. In seeds, PAH1/2 disruption minimally affected TAG content but significantly reduced LD size. Fatty acid feeding experiments demonstrated impaired LD formation and increased lipotoxicity in pah1/2 leaves and seedlings. Furthermore, knockout of PAH1/2 in mutants with enhanced fatty acid flux through phosphatidylcholine (PC) led to severe reductions in leaf TAG levels, despite increases in TAG synthesis rates, indicating accelerated TAG turnover. Phosphatidic acid, free fatty acids, and PC accumulated, leading to massive proliferation of endoplasmic reticulum membranes and severe growth and developmental defects. These findings demonstrate evolutionarily conserved roles for PAH1/2 in LD biogenesis, membrane lipid homeostasis, and cellular protection against lipotoxicity, particularly under conditions of elevated fatty acid flux.

59 BASIC BIOLOGICAL SCIENCES↗

LibraryX: A Framework for Cross-Library-Call Optimization

Scientific applications utilize performance libraries as a software engineering concept: these libraries encapsulate important and well-understood (mathematical) operations, allow for reuse, and are implemented and tuned by experts. Domain scientists then implement complex algorithms based on these domainspecific libraries. While individual library calls are optimized, larger performance gains across sequences of calls—sometimes spanning multiple libraries—are often unrealized, forcing a trade-off between performance and implementation complexity.To overcome this issue, we propose LibraryX, an approach and a system that allows for cross-library-call optimization even when library calls stem from multiple performance libraries. LibraryX annotates library calls with semantic information and optimizes entire directed acyclic graphs (DAGs) of calls dynamically using the SPIRAL code generation system. We demonstrate its effectiveness across a range of memory bound workloads, achieving significant speedups on Nvidia, AMD, and Intel accelerators compared to code using native libraries without cross-call optimization.

Rao, Sanil [Carnegie Mellon University,Department ↗

iDDS: intelligent distributed dispatch and scheduling for workflow orchestration

The intelligent distributed dispatch and scheduling (iDDS) service is a versatile workflow orchestration system designed for large-scale, distributed scientific computing. iDDS extends traditional workload and data management by integrating data-aware execution, conditional logic, and programmable workflows, enabling automation of complex and dynamic processing pipelines. Originally developed for the ATLAS experiment at the large hadron collider, iDDS has evolved into an experiment-agnostic platform that supports both template-driven workflows and a Function-as-a-Task model for Python-based orchestration. This paper presents the architecture and core components of iDDS, highlighting its scalability, modular message-driven design, and integration with systems such as PanDA and Rucio. We demonstrate its versatility through real-world use cases: fine-grained tape resource optimization for ATLAS, orchestration of large Directed Acyclic Graph (DAG) workflows for the Rubin Observatory, distributed hyperparameter optimization for machine learning applications, active learning for physics analyses, and AI-assisted detector design at the electron–ion collider. By unifying workload scheduling, data movement, and adaptive decision-making, iDDS reduces operational overhead and enables reproducible, high-throughput workflows across heterogeneous infrastructures. We conclude with current challenges and future directions, including interactive, cloud-native, and serverless workflow support.

97 MATHEMATICS AND COMPUTING↗

Deeplynx Airflow Provider Package

The DeepLynx Airflow Provider Package is a python package used to interact with the data warehouse DeepLynx when using the workflow orchestration tool Apache Airflow. This python package is packaged together using the airflow package standard so that it can be easily installed and used in any Apache Airflow environment. This package is meant to encapsulate the DeepLynx API for use in Airflow so that any interactions with DeepLynx that a user may want to use in their Airflow workflow can be easily accomplished using this provider package. This allows us to develop, implement, and test our DeepLynx-Airflow interactions in one provider package repository, and then easily install and use this package in any airflow instance. This DeepLynx Airflow Provider Package will be used extensively by the DeepLynx DAG repository.

Cavaluzzi, JackM↗

interactEM v1.0

An interactive, container-based workflow tool for creating and spawning directed acyclic graphs (DAGs) of operators in a distributed environment. It has a microservices architecture, and flow-based programming model. Current tools like this do not enable streaming of data directly between operators.

Welborn, Sam [Lawrence Berkeley National Laborator↗

Batched sparse direct solver design and evaluation in SuperLU_DIST

Over the course of interactions with various application teams, the need for batched sparse linear algebra functions has emerged in order to make more efficient use of the GPUs for many small and sparse linear algebra problems. In this paper, we present our recent work on a batched sparse direct solver for GPUs. The sparse LU factorization is computed by the levels of the elimination tree, leveraging the batched dense operations at each level and a new batched Scatter GPU kernel. The sparse triangular solve is computed by the level sets of the directed acyclic graph (DAG) of the triangular matrix. Batched operations overcome the large overhead associated with launching many small kernels. For medium sized matrix batches with not-so-small bandwidth, using an NVIDIA A100 GPU, our new batched sparse direct solver is orders of magnitude faster than a batched banded solver and uses less than one-tenth of the memory.

Boukaram, Wajih↗

Journey to Time-Variable Moment Tensors through Inversion of Acoustic and Seismoacoustic Data

We explore the capability of acoustic and seismoacoustic datasets to directly resolve a complex, time-variable source consisting of a buried mechanism, represented as a moment tensor, and a spall mechanism, represented as a vertical force at the surface. Traditionally, each component of a resolved moment tensor assumes one underlying source time function, which likely fails to capture the full evolution of a dynamic source, such as an explosion followed by slip on near-source joints or development of spallation. Specifically, we expand previous work to resolve a time-variable moment tensor using single-modality and joint-modality inversion frameworks through analysis of infrasound and seismoacoustic data recorded as part of the Source Physics Experiment Phase II: Dry Alluvium Geology (DAG). We investigate the impact of including signals from seismic-to-air coupling that are local to each infrasound sensor in comparison to mainly atmosphere-propagating acoustic signals, which occur from coupling of the wavefield from the subsurface to the atmosphere directly above the source. Additionally, we assess the ability of our inversion algorithm to fit observed infrasound data using a variety of time-variable source mechanisms. First, we consider the buried moment tensor source alone, which assumes that the determined Green’s functions incorporate effects from spallation or that the impact from spallation is minimal. Second, we examine the estimated buried moment tensor and vertical surface spallation as terms that must both be resolved in the inversion. Third, we assess the ability for an estimated vertical surface spallation source to fit the acoustic data on its own. Finally, we compare results from the joint inversion of both seismic geophone and infrasound acoustic data for the buried-only source compared to buried and spallation sources. Our results are a preliminary investigation into the applications of the inversion technique to recorded datasets and show the technique has limited capabilities using acoustic data alone. Instead, this method shows promise for seismic and seismoacoustic datasets to resolve the time-variable mechanisms of a buried source.

47 OTHER INSTRUMENTATION↗

CAMEO: A Co-design Architecture for Multi-objective Energy System Optimization (Project Report)

CAMEO (Codesign Architecture for Multi-objective Energy System Optimization) is a modular workflow management framework that abstracts co-design problems as Directed Acyclic Graphs (DAG). The framework employs JSON-based workflow specifications that enable systematic decomposition of complex optimization problems into reusable, interchangeable components including data loaders, scenario generators, optimization solvers, and result summarizers.

97 MATHEMATICS AND COMPUTING↗