Search NASA⌕ Search

SEARCH · Search NASA

Results for “Model Based System Engineering”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Knowledge Graph for End-to-End Traceability of an Integrated Human-Earth System Model

Integrated human-Earth system models inform energy-water-land system dynamics and policies, yet their results are difficult to trace through input-data, model structure, scenario configurations, and solved outputs. Because this information is siloed across disconnected artifacts, process-based IAMs have historically lacked a unified, queryable representation. Such lack of traceability prevents researchers from systematically isolating the multi-sector drivers of complex outcomes (such as tracing water-scarcity results back to distant energy-system dynamics) or conducting holistic uncertainty attribution across hundreds of interacting parameters. To address this concern, our work documents the software engineering process of a knowledge graph that unifies these four layers for the Global Change Analysis Model (GCAM-USA_Reference scenario, GCAM v9.1). The graph was built as a relational property graph in DuckDB from the run’s own artifacts: the input-preparation dependency map (gcamdata chunk map), the model’s XML input files, the run configuration, and the results database (BaseX), successfully mapping the model’s declared structure. The resulting graph comprises 204,321 nodes and 1,687,814 edges across 16 node types and 15 edge types, with approximately 16.3 million time-series values stored separately to maintain structural efficiency. To ensure representation fidelity, every edge carries an epistemic-status annotation recording the warrant for the relationship (structural, provenance, dependency, or model-derived), and a machine-readable provenance ledger classifying the origin of every schema element. Evaluation against a fixed five-benchmark suite with locked baselines reports zero structural orphans, zero dangling edge endpoints, and 100% of output-producing technologies traceable to raw input files. Two interactive interfaces present the graph, including a serverless browser application built on DuckDB-Wasm. By establishing the first end-to-end provenance framework for an IAM, this work enables researchers and scientists to systematically audit complex policy scenarios, debug model structures, and trace policy-relevant outputs to their data origins in real time.

Artifical Intelligence↗

Bridging Cloud and Edge Computing at NREL Using CONNECT: Cloud Optimized Networking for Next-Gen Edge Computing Technologies [Slides]

CONNECT is an innovative on-premise hardware and software solution that integrates edge and cloud computing infrastructure at NREL. Built on the AWS Greengrass middleware and leveraging the MQTT protocol, CONNECT enables real-time data streaming from IoT devices and gateways to both cloud and local services, empowering researchers to rapidly capture, analyze, and act upon edge-generated data while leveraging cloud capabilities. The platform addresses research infrastructure challenges by providing a pre-approved platform which is already configured with the correct networking and cybersecurity baselines thus eliminating procurement delays and enabling on-demand availability. CONNECT's hybrid architecture efficiently manages burstable workloads, allowing research teams to dynamically scale computational capacity, handle peak data loads, and reduce operational bottlenecks. Advanced capabilities include built-in GPU support for executing machine learning models which enables low-latency inference at the edge from models trained in the cloud. This architecture supports real-time analytics and filtering, providing a mechanism to allow only transmitting and processing high-value data. Cloud-based configuration management permits engineers to manage on-premise systems remotely, optimizing operational efficiency. By bridging edge and cloud computing, CONNECT provides NREL researchers with a flexible, scalable platform that accelerates scientific discovery while maintaining robust security and performance standards.

97 MATHEMATICS AND COMPUTING↗

IMAGINE BioSecurity: Mesocosm-Based Methods to Evaluate Biocontainment Strategies and Impact of Industrial Microbes Upon Native Ecosystems

Project Goals: The Integrative Modeling and Genome-scale Engineering for Biosystems Security (IMAGINE BioSecurity) SFA project seeks to establish an understanding of the behavior of engineered microbes in controlled versus environmental conditions to predictively devise new strategies for responding to biological escape. To this end, the IMAGINE Team has established a plant-soil mesocosm platform to track and quantify the fate of industrial microbes in environmental systems and assess the efficacy of biocontainment constraints upon genetically engineered microbe escape frequency and the impact of industrial microbes upon native ecological microbiomes. Abstract Text: Genetically modified industrial production microbes and their associated bioproducts have emerged as an integral component of a sustainable bioeconomy. However, the rapid development of these innovative technologies raises biosecurity concerns, namely, the risk of environmental escape. Thus, the realization of a bioeconomy hinges not only on the development and deployment of microbial production hosts, but also on the development of secure biosystems and biocontainment designs. Current laboratory-based biocontainment testing systems do not accurately reflect complexities found in natural environments, necessitating an environmentally relevant analysis pipeline that allows for the detection of rare escapees, the effect of associated bio-products, and the impact on native ecologies. To this end, we have developed an approach that utilizes soil mesocosms and integrated systems analyses to evaluate the efficacy of novel biocontainment strategies and to assess the impact of production systems upon terrestrial microbiome dynamics. We demonstrate the utility of this approach by modeling a contamination with industrial microbial chasses versus their biocontained counterparts. Here we demonstrate the broad utility of this system by highlighting findings from both strains of Saccharomyces cerevisiae that are contained with an inducible toxin anti-toxin system, and stains of Escherichia coli that are contained via genomic recoding. The resultant data demonstrate that this system has broad utility across diverse microbial chassis and biocontainment strategies, enables us to track the fate of our contaminating microbe with high sensitivity in the soil, as well as monitor broader impacts of the perturbation on the underlying soil system. The findings presented here support the use of this mesocosm-based approach to assess the environmental impact of industrial microbes and to validate biocontainment strategies.

BASIC BIOLOGICAL SCIENCES,INORGANIC, ORGANIC, PHYS↗

An investigation on machine learning predictive accuracy improvement and uncertainty reduction using VAE-based data augmentation

The confluence of ultrafast computers with large memory, rapid progress in Machine Learning (ML) algorithms, and the availability of large datasets place multiple engineering fields at the threshold of dramatic progress. However, a unique challenge in nuclear engineering is data scarcity because experimentation on nuclear systems is usually more expensive and time-consuming than most other disciplines. One potential way to resolve the data scarcity issue is deep generative learning, which uses certain ML models to learn the underlying distribution of existing data and generate synthetic samples that resemble the real data. In this way, one can significantly expand the dataset to train more accurate predictive ML models. In this study, our objective is to evaluate the effectiveness of data augmentation using variational autoencoder (VAE)-based deep generative models. We investigated whether the data augmentation leads to improved accuracy in the predictions of a deep neural network (DNN) model trained using the augmented data. Additionally, the DNN prediction uncertainties are quantified using Bayesian Neural Networks (BNN) and conformal prediction (CP) to assess the impact on predictive uncertainty reduction. To test the proposed methodology, we used TRACE simulations of steady-state void fraction data based on the NUPEC Boiling Water Reactor Full-size Fine-mesh Bundle Test (BFBT) benchmark. Here, we found that augmenting the training dataset using VAEs has improved the DNN model’s predictive accuracy, improved the prediction confidence intervals, and reduced the prediction uncertainties.

Bayesian neural network↗

Algebraic Multigrid with Filtering: An Efficient Preconditioner for Interior Point Methods in Large-Scale Contact Mechanics Optimization

Large-scale contact mechanics simulations are crucial in many engineering fields such as structural design and manufacturing. In the frictionless case, contact can be modeled by minimizing an energy functional; however, these problems are often nonlinear, nonconvex, and increasingly difficult to solve as mesh resolution increases. In this work, we employ a Newton-based interior-point (IP) filter line-search method, an effective approach for large-scale constrained optimization. While this method converges rapidly, each iteration requires solving a large saddle-point linear system that becomes ill-conditioned as the optimization process converges, largely due to IP treatment of the contact constraints. Such ill-conditioning can hinder solver scalability and increase iteration counts with mesh refinement. Here, to address this, we introduce a novel preconditioner, algebraic multigrid with filtering (AMGF), tailored to the Schur complement of the saddle-point system. Building on the classical AMG solver, commonly used for elasticity, we augment it with a specialized subspace correction that filters near null space components introduced by contact interface constraints. Through theoretical analysis and numerical experiments on a range of linear and nonlinear contact problems, we demonstrate that the proposed solver achieves mesh independent convergence and maintains robustness against the ill-conditioning that notoriously plagues IP methods. These results indicate that AMGF makes contact mechanics simulations more tractable and broadens the applicability of Newton-based IP methods in challenging engineering scenarios. More broadly, AMGF is well suited for problems, optimization or otherwise, where solver performance is limited by a low-dimensional subspace, such as those arising from localized constraints, interface conditions, or model heterogeneities. This makes the method widely applicable beyond contact mechanics and constrained optimization.

Mathematics and Computing↗

Conformational Dynamics and Catalytic Backups in a Hyper-thermostable Engineered Archaeal Protein Tyrosine Phosphatase

Protein tyrosine phosphatases (PTPs) are a family of enzymes that play important roles in regulating cellular signaling pathways. The activity of these enzymes is regulated by the motion of a catalytic loop that places a critical conserved aspartic acid side chain into the active site for acid–base catalysis upon loop closure. These enzymes also have a conserved phosphate-binding loop that is typically highly rigid and forms a well-defined anion-binding nest. The intimate links between loop dynamics and chemistry in these enzymes make PTPs an excellent model system for understanding the role of loop dynamics in protein function and evolution. In this context, archaeal PTPs, which have often evolved in extremophilic organisms, are highly understudied, despite their unusual biophysical properties. We present here an engineered chimeric PTP (ShufPTP) generated by shuffling the amino acid sequence of five extant hyperthermophilic archaeal PTPs. Despite ShufPTP’s high sequence similarity to its natural counterparts, it presents a suite of unique properties, including high flexibility of the phosphate binding P-loop, facile oxidation of the active-site cysteine, mechanistic promiscuity, and, most notably, hyperthermostability, with a denaturation temperature likely >130 °C (>8 °C higher than the highest recorded growth temperature of any archaeal strain). Our combined structural, biochemical, biophysical, and computational analysis provides insight both into how small steps in evolutionary space can radically modulate the biophysical properties of an enzyme and showcases the tremendous potential of archaeal enzymes for biotechnology, to generate novel enzymes capable of operating under extreme conditions.

archaea↗

SigTime: Learning and Visually Explaining Time Series Signatures

Understanding and distinguishing temporal patterns in time series data is essential for scientific discovery and decision-making. For example, in biomedical research, uncovering meaningful patterns in physiological signals can improve diagnosis, risk assessment, and patient outcomes. However, existing methods for time series pattern discovery face major challenges, including high computational complexity, limited interpretability, and difficulty in capturing meaningful temporal structures. Here, to address these gaps, we introduce a novel learning framework that jointly trains two Transformer models using complementary time series representations: shapelet-based representations to capture localized temporal structures and traditional feature engineering to encode statistical properties. The learned shapelets serve as interpretable signatures that differentiate time series across classification labels. Additionally, we develop a visual analytics system—SigTime—with coordinated views to facilitate exploration of time series signatures from multiple perspectives, aiding in useful insights generation. We quantitatively evaluate our learning framework on eight publicly available datasets and one proprietary clinical dataset. Additionally, we demonstrate the effectiveness of our system through two usage scenarios along with the domain experts: one involving public ECG data and the other focused on preterm labor analysis.

97 MATHEMATICS AND COMPUTING↗

Enhancing Operational Safety via Agentic Dialogue Hazard Identification Analysis

Operational safety in high-stakes domains such as industrial process control, autonomous, and safety-critical systems demand reliable hazard identification. While large language models (LLMs) have shown promise in automating safety analysis tasks, single-turn, monolithic inference is brittle: it lacks the self-correction, deliberation, and contextual refinement that safety engineers apply iteratively. In this paper, we introduce HAZDIAL, a framework that investigates whether structured agentic dialogue (multi-agent, multi-turn interactions) improves the quality of NLP-based hazard identification over single-pass baselines. We systematically compare two dialogue modalities: adversarial debate and constructive discussion, and propose an genetic algorithm-based agentic interaction optimization. We evaluate all configurations against a curated golden dataset using standard classification metrics (accuracy, precision, recall, F1) and a novel dialogue metrics. This work advances the intersection of dialogue systems, multi-agent reasoning, and AI safety, providing empirical evidence for dialogue-driven hazard analysis.

Das, Sanjay [ORNL] (ORCID:0009000542591915)↗

Development of a scintillator based fast-ion loss detector for the Wendelstein 7-X stellarator

A new scintillator based fast-ion loss detector (FILD) system has been designed for the Wendelstein 7-X (W7-X) stellarator. The mechanical design of the system is presented here along with engineering analyses of the system. This includes an assessment of the structural loads which also considers electromagnetic forces calculated for possible disruption scenarios as well as a mode frequency analysis of the system. Furthermore, an analysis of the thermal characteristics of the actively cooled probe head under prescribed steady state and transient heat loads is presented. Finally, the optical relay system is described and its properties used to forward model synthetic camera signals using the FILDSIM code and based on theoretical fast-ion losses incident on the probe head calculated using the ASCOT5 code.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

SetGo: Metadata Readiness for Scientific AI Datasets

Scientific datasets intended for AI use require both computational readiness for model training and metadata readiness for discovery, sharing, and reuse. The Readiness Engine for Data Integration (REDI) addresses computational readiness, but no corresponding tool evaluates whether a dataset’s metadata are sufficiently complete, governed, and standards-compliant for publication and agent-based consumption. Existing FAIR assessors operate only on published repository records, and no single system covers FAIR compliance, licensing, provenance, governance, reproducibility, and catalog readiness together. We present SetGo, an open-source Python toolkit that assesses and repairs metadata readiness across these six dimensions before a dataset is published or archived. Applied to four scientific corpora, SetGo surfaces deficiencies that general-purpose tools do not detect: ERA5 climate metadata scores 4% on ACDD 1.3 compliance; materials datasets fail OPTIMADE species-definition requirements; and PDB-derived proteomics data carries licensing terms incompatible with standard SPDX identifiers. Guided enrichment raises overall FAIR scores from 52–57% to 81–91%, and a single setgo publish command pushes to Hugging Face Hub, CKAN, or OpenMetadata with ML Commons Croissant 1.0 metadata sidecars. To support interactive and automated workflows, SetGo integrates with coding agents powered by large language models (LLMs) through a /setgo skill that enables natural-language execution of the full assess–enrich–publish loop, with user involvement limited to supplying missing metadata values.

Wilkinson, Sean [ORNL] (ORCID:0000000214437479)↗

Producing multiple chemicals through biological upcycling of waste poly(ethylene terephthalate)

Poly(ethylene terephthalate) (PET) waste is of low degradability in nature, and its mismanagement threatens numerous ecosystems. To combat the accumulation of waste PET in the biosphere, PET bio-upcycling, which integrates chemical pretreatment to produce PET-derived monomers with their microbial conversion into value-added products, has shown promise. The recently discovered Rhodococcus jostii strain PET (RPET) can metabolically degrade terephthalic acid (TPA) and ethylene glycol (EG) as sole carbon sources, and it has been developed into a microbial chassis for PET upcycling. However, the scarcity of synthetic biology tools, specifically designed for the non-model microbe RPET, limits the development of a microbial cell factory for expanding the repertoire of bioproducts from post-consumer PET. Herein, we describe the development of potent genetic tools for RPET, including (1) two inducible and titratable expression systems for tunable gene expression and (2) Serine Integrase-based Recombinational Tools (SIRT) for genome editing. Using these tools, we systematically engineer the RPET strain to ultimately establish microbial supply chains for producing multiple chemicals, including lycopene, lipids, and succinate, from post-consumer PET waste bottles, achieving the highest titer of lycopene ever reported thus far in RPET (i.e., 22.6 mg/L of lycopene, approximately 10,000-fold higher than that of the wild-type strain). Furthermore, this work highlights the great potential of plastic upcycling as a generalizable means of sustainable production of diverse chemicals.

36 MATERIALS SCIENCE↗

Pressure Gain, Stability, and Operability of Methane/Syngas Based RDEs Under Steady and Transient Conditions (Final Project Report)

The scope of this work addresses key issues associated with losses associated with the detonation wave and other processes internal to the RDE operation, as well as it develops modeling tools for the evaluation of these losses and exhaust emissions in RDEs. The main challenge in studying RDEs is that RDE performance is highly reliant on the specifics of the design so much so that simple/canonical systems alone cannot provide useful engineering information, but practical RDE designs are sufficiently complex and involve extreme operational environments that detailed access either experimentally (laser diagnostics, for instance) or computationally (direct numerical simulations) are as yet to become practical. To overcome this challenge, we have conducted a combined experimental/simulation/analytical study investigating key phenomena that control the characteristics of operation of RDEs. As a result, the study has developed tools and methods that can be used to evaluate performance and design approaches using reduced-physics models, with the assumptions validated using detailed simulations, and the model prediction tested using experimental observations. The specific objectives of the research were: (1) Develop and demonstrate a low-loss fully axial injection concept, taking advantage of stratification effects to alter the detonation structure and position the wave favorably within the combustor; (2) Obtain stability and operability characteristics of an RDE across operating conditions to aid in the development of operability and performance rules for the operations of other systems; and (3) Develop quantitative metrics for performance gain as well as quantitative description of the loss mechanisms through a combination of diagnostics development, reduced-order modeling, and detailed simulations. The work conducted here has made contribution on design of low-loss inlets that has broad application within the power generation industry for use with pressure gain combustion. The operability and stability of different designs, while focusing on axial air inlet designs, has been analyzed. The effect of nozzle and injection conditions was studied. Models and simulations of exhaust emissions, focusing on NOx emission has been developed and used to investigate how operation of the RDE affect NOx production using Lagrangian analysis of RDE simulations. This work has built on previous programs, with the goal of further understanding operation of RDEs and elevate the readiness of design consideration. In addition, a suite of diagnostic and modeling tools have been developed to obtain quantitative metrics on performance based on measurements, which can be readily transferred to other experimental configurations.

08 HYDROGEN↗

Connecting Minds: AI Use Cases to Bridge Power Systems and Large Language Models for Practical Applications

Recent advances in artificial intelligence (AI) and development of large language models (LLMs) present the opportunity to develop a new generation of power systems applications. In contrast with early power system AI applications based on structured numerical data, LLMs offer unique capabilities to perform logical reasoning using text documents, unstructured data, and application programming interface (API) calls to computational software. This paper seeks to bridge the knowledge gap between power systems engineers and LLM developers through a crosscutting explanation of use cases, characteristics, requirements, practical considerations from the perspectives of both LLM capabilities and industry needs. Specific focus is given to applications that can be realistically deployed by electric utilities. After introducing the architecture of LLMs and unique challenges of the power systems domain, this paper proposes twenty representative LLM applications grouped into categories of 1) power system operations, 2) asset management, 3) system planning and analytics, and 4) energy management and protection systems. Five use cases are presented within each category with descriptions of the motivation, objectives, approaches, example inputs / outputs, and benefits of each use case.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Graph-Based Attention Mechanisms for Solving the AC Optimal Power Flow Problem in Electrical Power Networks

With the increasing complexity and data availability in modern power systems, learning-based approaches to AC Optimal Power Flow (AC OPF) have garnered significant attention. In particular, the structure of smart grids lends itself naturally to graph-based representations, where Graph Neural Networks (GNNs) can capture spatial and relational dependencies. This paper investigates attention-based GNN architectures tailored to heterogeneous graph representations of electric grids. We evaluate two major paradigms: relational attention, which distinguishes between edge types during message passing, and meta-path attention, which captures high-level semantics through multi-hop, typed paths. Using a large corpus of public AC OPF scenarios, we benchmark representative models of each type of attention. Our results demonstrate the benefits of heterogeneous attention-based models in accurately capturing grid dynamics; heterogeneous attention models achieve superior performance in both standard and perturbed settings. The findings highlight the importance of semantic-aware architectures for improving prediction robustness and interpretability in power system applications.

Trigui, Ali [Qubit Engineering Inc.]↗

Feedback and Oscillations: Constructing Feedback Systems for Root Cause Analysis of Oscillations in Power Grids

Dynamic phenomena linked to inverter-based resources (IBRs) have gained global attention. Several IBR-induced dynamics have caused bulk power system-connected wind or solar power plants to trip, and some have even led to widespread outages. In addition, many oscillations have been observed involving IBR power plants. In 2023, the IEEE Power & Energy Society (PES) IBR Subsynchronous Oscillations (SSO) task force published a journal article, “Real-World Subsynchronous Oscillation Events in Power Grids With High Penetrations of Inverter-Based Resources,” in which 19 IBR oscillation events were examined for their causation. Earlier in 2020, another PES task force article, “Definition and Classification of Power System Stability-Revisited & Extended,” authored by prominent academics, introduced converter-driven stability as a new category of stability. The international power grid industry community also took action by publishing the CIGRE Green Book, Power System Dynamic Modelling and Analysis in Evolving Networks (led by Babak Badrzadeh and Zia Emin) in 2024. In August 2024, the Energy Systems Integration Group (ESIG) released a practical guide led by Nick Miller, “Diagnosis and Mitigation of Observed Oscillations in IBR-Dominant Power System: A Practical Guide.” The goal of the guide is to assist practicing engineers in making initial judgments and conducting detailed analyses about oscillations. Finally, when addressing the classification of stability and oscillations, the guide emphasizes a causality-based taxonomy for grouping, such as voltage control-induced oscillations, synchronization-induced oscillations, and frequency or active power control-induced oscillations.

Fan, Lingling [Univ. of South Florida, Tampa, FL (↗

Review and Modeling of Integrated Energy Systems with Nuclear Reactor Coupled Desalination and District Heating

Detailed reviews of a past advanced nuclear reactor based integrated energy system, as well as other nuclear reactor and fossil fuel based integrated energy systems have been performed for this work. Review of the utilization of heat from nuclear reactors for various applications and cogeneration has been done. The heat can be utilized by extraction of the steam from the turbine while the steam is still at a desired temperature. While use of nuclear process heat for district heating in countries like Finland, France, China, Poland, and elsewhere is discussed, more focus of the review has been given on nuclear desalination processes. Integrated energy systems (IES) where distinct types of reactors like PWR, BWR, sodium cooled fast reactor, heavy water reactor and other advanced reactors are coupled with various nuclear desalination processes like multi-effect distillation (MED), multi-stage flashing (MSF) and reverse osmosis (RO) methods have been discussed. The nuclear desalination plant at Aktau has been discussed in more detail due to its decades of successful operation. The IES of the Aktau plant coupled with 5-effect MED desalination plant has been taken as a reference for modeling the Open Modelica (OM) based IES of this work. Here, the OM IES model shows good agreement with the MED plant output of Aktau and can be extended for future applications of IES.

42 ENGINEERING↗

Best practices in software development for robust and reproducible geoscientific models based on insights from the Global Carbon Budget's dynamic vegetation models

Computational models play an increasingly vital role in scientific research by enabling the numerical simulation of complex processes. Such models are also fundamental in geosciences. For instance, they offer critical insights into the impacts of global change on the Earth system today and in the future. Beyond their value as research tools, models are also software products and should therefore adhere to certain established software engineering standards. However, scientists are rarely trained as software developers, which can lead to potential deficiencies in software quality like unreadable, inefficient, or erroneous code. The complexity of models, coupled with their integration into broader workflows, also often makes it challenging to reproduce results, evaluate processes, and build upon them. In this paper, we review the state and current practices of the development processes of the state-of-the-art land surface models used by the Global Carbon Budget. We combine the experience of modelers from the respective research groups with the expertise of software engineers from tech companies to outline key principles and tools for improving software quality in research. We explore four main areas: (1) model testing and validation, (2) scientific, technical, and user documentation, (3) version control, continuous integration, and code review, and (4) the portability and reproducibility of workflows. Our review reveals that while modeling communities are incorporating many best practices, significant room for improvement remains in areas such as automated testing, automated documentation, and reproducibility. Therefore, we here identify and promote essential software engineering practices, including numerous examples of practices from within the community that can serve as guidelines for other models and could help streamline processes across the entire community. We conclude with an open-source example implementation of these principles, demonstrating portable and reproducible data flows, a continuous integration setup, and web-based visualizations. This example may serve as a practical resource for model developers, users, and all scientists engaged in scientific programming.

Gregor, Konstantin [Technical Univ. of Munich (Ger↗

Developing multi-gene CRISPRa/i programs to accelerate DBTL cycles in ABF hosts engineered for chemical production

This project developed and implemented a modular CRISPR activation and interference (CRISPRa/i) platform to accelerate strain optimization and pathway development for industrially relevant microbial hosts. By integrating multiplexed transcriptional perturbation tools with data-driven Design–Build–Test–Learn (DBTL) workflows, the team achieved reductions in cycle time and enhanced production of industrial aromatics, particularly 4-aminocinnamic acid (4-ACA), in Pseudomonas putida. Key accomplishments included: ● Development of a robust, tunable CRISPRa/i system in P. putida that enabled efficient multi-target gene regulation via guide RNA (gRNA) programs ● Completion of two full DBTL cycles, guided by machine learning (ML) models trained on transcriptomic and performance data, reducing engineering time by over 30% ● Optimization of multi-gene regulatory programs to balance expression of host and pathway modules, improve 4-ACA titers, and resolve metabolic bottlenecks ● Demonstration of system portability through a limited proof-of-concept extension in Acinetobacter baylyi, underscoring the generalizability of the approach ● Evaluation of strain performance on lignocellulosic biomass-derived substrates, demonstrating the feasibility of converting renewable carbon into aromatic building blocks These results illustrate the feasibility of applying ML-guided CRISPRa/i perturbation strategies to accelerate strain development in complex microbial systems. The resulting tools and datasets contribute to DOE objectives by improving platform predictability, reducing development costs, and enabling broader access to sustainable, economically viable bioproduction technologies.

09 BIOMASS FUELS↗