Search NASA⌕ Search

SEARCH · Search NASA

Results for “data platform”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 379 records · Page 21

Understanding Aitken Mode Aerosol Variability over the Southern Ocean and Antarctica: Insights from Cloud Condensation Nuclei Data

Aitken mode aerosol particles play an important role influencing cloud properties and sustenance, acting as a reservoir of potential cloud condensation nuclei against precipitation scavenging. However, there is limited data on Aitken mode aerosols. In this study, we develop a method to estimate Aitken mode aerosol concentrations and size distribution using cloud condensation nuclei measurements (CCN) and κ-Köhler theory. The performance of this method is evaluated using scanning mobility particle sizer (SMPS) data from recent field campaigns to demonstrate its skills and applicability. The method reasonably estimates Aitken- and accumulation-mode aerosol concentrations, achieving correlations of 0.7–0.9 with only modest biases (mean fractional bias within ±23% for Aitken mode and ±34% for accumulation-mode). This method is further applied to measurements collected over the Southern Ocean and Antarctica in recent years from multiple platforms, including ground sites, aircraft, and ships, to derive Aitken and accumulation-mode aerosol concentrations. Using the derived data, we examine the seasonal cycle, latitudinal variations, and vertical distribution of aerosols. Aitken mode aerosol concentrations are elevated over the Southern Ocean and Antarctica during the austral summer similar to the accumulation mode. In the austral summer, the free troposphere has more Aitken mode aerosols and fewer accumulation mode aerosols than the boundary layer, and thus likely serves as an important source of cloud-forming aerosol while also diluting the accumulation mode.

Kang, Litai [University of Washington] (ORCID:0000↗

Modification and analysis of context-specific genome-scale metabolic models: methane-utilizing microbial chassis as a case study

ABSTRACT Context-specific genome-scale model (CS-GSM) reconstruction is becoming an efficient strategy for integrating and cross-comparing experimental multi-scale data to explore the relationship between cellular genotypes, facilitating fundamental or applied research discoveries. However, the application of CS modeling for non-conventional microbes is still challenging. Here, we present a graphical user interface that integrates COBRApy, EscherPy, and RIPTiDe, Python-based tools within the BioUML platform, and streamlines the reconstruction and interrogation of the CS genome-scale metabolic frameworks via Jupyter Notebook. The approach was tested using -omics data collected for Methylotuvimicrobium alcaliphilum 20Z R , a prominent microbial chassis for methane capturing and valorization. We optimized the previously reconstructed whole genome-scale metabolic network by adjusting the flux distribution using gene expression data. The outputs of the automatically reconstructed CS metabolic network were comparable to manually optimized i IA409 models for Ca-growth conditions. However, the CS model questions the reversibility of the phosphoketolase pathway and suggests higher flux via primary oxidation pathways. The model also highlighted unresolved carbon partitioning between assimilatory and catabolic pathways at the formaldehyde-formate node. Only a very few genes and only one enzyme with a predicted function in C1 metabolism, a homolog of the formaldehyde oxidation enzyme ( fae1-2 ), showed a significant change in expression in La-growth conditions. The CS-GSM predictions agreed with the experimental measurements under the assumption that the Fae1-2 is a part of the tetrahydrofolate-linked pathway. The cellular roles of the tungsten (W)-dependent formate dehydrogenase ( fdhAB ) and fae homologs ( fae1-2 and fae3 ) were investigated via mutagenesis. The phenotype of the f dhAB mutant followed the model prediction. Furthermore, a more significant reduction of the biomass yield was observed during growth in La-supplemented media, confirming a higher flux through formate. M. alcaliphilum 20Z R mutants lacking fae1-2 did not display any significant defects in methane or methanol-dependent growth. However, contrary to fae1, the fae1-2 homolog failed to restore the formaldehyde-activating enzyme function in complementation tests. Overall, the presented data suggest that the developed computational workflow supports the reconstruction and validation of CS-GSM networks of non-model microbes. IMPORTANCE The interrogation of various types of data is a routine strategy to explore the relationship between genotype and phenotype. An efficient approach for integrating and cross-comparing experimental multi-scale data in the context of whole-genome-based metabolic network reconstruction becomes a powerful tool that facilitates fundamental and applied research discoveries. The present study describes the reconstruction of a context-specific (CS) model for the methane-utilizing bacterium, Methylotuvimicrobium alcaliphilum 20Z R . M. alcaliphilum 20Z R is becoming an attractive microbial platform for the production of biofuels, chemicals, pharmaceuticals, and bio-sorbents for capturing atmospheric methane. We demonstrate that this pipeline can help reconstruct metabolic models that are similar to manually curated networks. Furthermore, the model is able to highlight previously overlooked pathways, thus advancing fundamental knowledge of non-model microbial systems or promoting their development toward biotechnological or environmental implementations.

Kulyashov, M. A.↗

Modularization of EDGE Workflows Using Nextflow: Improving the Efficiency and Maintainability of Bioinformatics Software

EDGE is a bioinformatics platform developed in 2016 by researchers at Los Alamos National Laboratory (LANL) to facilitate the analysis of next-generation sequencing data by researchers with varying levels of experience in bioinformatics (Li et al., 2017). Users with single-end, paired-end or long-read sequencing data can provide their reads as input to EDGE and select the combination of workflows to run that are most useful for their research (e.g., quality control of reads, genome assembly, or the taxonomic classification of input reads). Table 1 summarizes the modules available in EDGE. EDGE is available as a web platform at https://edgebioinformatics.org, as installable source code maintained on GitHub under a GPLv3 license, and as a publicly hosted Docker image.

59 BASIC BIOLOGICAL SCIENCES↗

Hydrodynamic characterization of the coastal pioneer array ocean observing system

Ocean observation buoys require relatively small amounts of power, yet traditionally necessitate costly resupply trips for battery replacement. With the offshore location of the buoys and small power requirements, wave energy may be an effective solution for providing consistent and reliable power to support the buoy instrumentation. The US National Science Foundation Ocean Observatories Initiative (OOI) includes arrays of point absorber-like buoy systems used for ocean observation that have been deployed at multiple locations including the Southern Mid-Atlantic Bight. A study is currently underway to design a pitch resonator wave energy converter to supplement existing renewable energy generation for powering observation instrumentation. This paper details field measurements from surface moorings of the OOI Coastal Pioneer Array, which informs the subsequent development of a numerical model for the moored observation system. The model is developed in Wave Energy Converter Simulator (WEC-Sim), which leverages the Simscape multibody solver within the MATLAB/Simulink framework and linear potential flow theory to simulate the hydrodynamic interactions and multibody dynamics in 6 degrees of freedom. Multiple tuning variables are considered to produce a model for the system that matches well with empirical data (about 8% error). In conclusion, the WEC-Sim model will serve as a platform for integrating the pitch resonator wave energy converter concept and deployment preparation (detailed design including power take-off and control systems, response evaluation, etc.).

hydrodynamic modeling↗

32 examples of LLM applications in materials science and chemistry: towards automation, assistants, agents, and accelerated scientific discovery

Abstract Large language models (LLMs) are reshaping many aspects of materials science and chemistry research, enabling advances in molecular property prediction, materials design, scientific automation, knowledge extraction, and more. Recent developments demonstrate that the latest class of models are able to integrate structured and unstructured data, assist in hypothesis generation, and streamline research workflows. To explore the frontier of LLM capabilities across the research lifecycle, we review applications of LLMs through 32 total projects developed during the second annual LLM hackathon for applications in materials science and chemistry, a global hybrid event. These projects spanned seven key research areas: (1) molecular and material property prediction, (2) molecular and material design, (3) automation and novel interfaces, (4) scientific communication and education, (5) research data management and automation, (6) hypothesis generation and evaluation, and (7) knowledge extraction and reasoning from the scientific literature. Collectively, these applications illustrate how LLMs serve as versatile predictive models, platforms for rapid prototyping of domain-specific tools, and much more. In particular, improvements in both open source and proprietary LLM performance through the addition of reasoning, additional training data, and new techniques have expanded effectiveness, particularly in low-data environments and interdisciplinary research. As LLMs continue to improve, their integration into scientific workflows presents both new opportunities and new challenges, requiring ongoing exploration, continued refinement, and further research to address reliability, interpretability, and reproducibility.

Computer Science↗

Machine learning and deep learning tools for the automated capture of cancer surveillance data

The National Cancer Institute and the Department of Energy strategic partnership applies advanced computing and predictive machine learning and deep learning models to automate the capture of information from unstructured clinical text for inclusion in cancer registries. Applications include extraction of key data elements from pathology reports, determination of whether a pathology or radiology report is related to cancer, extraction of relevant biomarker information, and identification of recurrence. With the growing complexity of cancer diagnosis and treatment, capturing essential information with purely manual methods is increasingly difficult. These new methods for applying advanced computational capabilities to automate data extraction represent an opportunity to close critical information gaps and create a nimble, flexible platform on which new information sources, such as genomics, can be added. This will ultimately provide a deeper understanding of the drivers of cancer and outcomes in the population and increase the timeliness of reporting. These advances will enable better understanding of how real-world patients are treated and the outcomes associated with those treatments in the context of our complex medical and social environment.

60 APPLIED LIFE SCIENCES↗

Supervisory Control and Data Acquisition for Electrochemical Separation Experimentation

The Python-based program is a laboratory automation tool designed to control and monitor electrochemical systems. The tool was developed for capacitive deionization (CDI) experiments, but it can be used for any system that requires controlled voltage or current segments and multi-parameter monitoring. The program integrates hardware components to run user-defined experimental parameters, providing operational control of a programmable power supply, peristaltic pump, and data acquisition devices. Currently, the program is structured with a workflow that includes an initialization (or pre-run) phase, a main loop, and a post-experiment stabilization (or post-run) phase. The initialization phase prepares and stabilizes the cell, ensuring that the electrodes and solution reach a baseline state before the experiment begins. The main loop consists of multiple voltage segments that repeat, controlling the experiment while recording key parameters such as time, voltage, current, pH, and conductivity. Finally, the post-experiment stabilization phase allows the system to stabilize after the experiment, returning the cell and solution to equilibrium conditions before ending the sequence. The program is designed with four variations, each tailored to different experimental needs. All variations include both the initialization and post-experiment stabilization stages, which run for a set amount of time, voltage, current, and flow rate before and after the main experiment block. The main loop runs for a set number of cycles, as defined by the user input, and each cycle is composed of 2 or 4 segments. The 4 program variations are described as follows: Program 1: The main program includes 2 segments. Each segment is defined to have a set duration, flow rate, voltage, and current. This program measures conductivity, flow rate, voltage, and current. Program 2: The main program expands Program 1 to include 4 segments. Each segment has a specified duration, flow rate, voltage, and current. Like Program 1, it measures conductivity, flow rate, voltage, and current. Program 3: The main program consists of 2 segments, each defined by time, flow rate, voltage, and current. In addition to conductivity, flow rate, voltage, and current, Program 3 collects pH and temperature data through a 4-channel data acquisition device. Program 4: This program independently controls two channels of a multi-channel power supply simultaneously. While conductivity can only be measured for one cell at a time, the dual-channel control makes it possible to operate two cells simultaneously under different voltage/current conditions. The main program includes 2 segments.For each program, all measurements are automatically logged and integrated into a single Excel output file. Data are displayed in numerical format and plotted, both in real time, to track system performance. A key feature of the program is its ability to synchronize all outputs so that every measurement shares a single timestamp, ensuring accurate alignment of voltage, current, pH, conductivity, and pH data.By combining hardware control, real-time monitoring, and unified data collection, this program significantly reduces manual workload and minimizes errors, making it a reliable platform for researchers, engineers, and laboratory technicians conducting CDI experiments, among other electrochemical tests.

Valentino, Lauren [Argonne National Laboratory (AN↗

Platform for Remote Deployment and Training for Enhanced Building Operation Practices (Building Re-Tuning and On-going Commissioning)

While a building’s energy usage is driven largely by its design and use, building operator behavior has a strong influence on its energy consumption. This project developed and piloted a specific, data-driven coaching methodology to help operators understand how they can adjust operations and/or affect no/low-cost repairs or upgrades to their specific building HVAC systems to reduce energy consumption. Named BuildingCoach, the operational optimization method used is based on the Building Re-tuning approach developed by the Pacific Northwest National Laboratory. A building operations analytics market has matured over the past decade, though its potential to affect energy-saving changes has not been fully realized. Training operators to understand the methods for operational optimization with the explicit approach of using building-system performance data is hypothesized to create a more effective, longer lasting result in building energy efficiency, and this strategy is the fundamental premise of this project. With the support of an Industry Advisory Board, the project succeeded in developing materials and recruiting for and delivering three pilot cohorts. Deliverables included twenty-two self-paced training modules (accessed via a Learning Management System) and a web-based platform that includes access to real-time building system data and a repository for building system documentation. The project set out to have 100 participants from 50 buildings in three pilot cohorts. In the end, there were 28 participants from 17 buildings, i.e., a significant shortfall. The first two pilot cohorts had only two buildings in each, and this was partially due to difficulties in deploying the Building Operator Coaching Solution (“the BOCS”), which is technology that extracts the data from the controls network and presents it as prescribed for coaching. In the third cohort, the project team deployed the BOCS successfully to 13 buildings, the methodology was piloted as intended, and numerous opportunities for optimization were identified. The BuildingCoach business plan charts a path to an economically sustainable effort. However, even with a licensing model captured in the final version of the business plan, the scalability is still limited to keeping less than 1,000 buildings affected by 2033. Even so, there are unexplored paths to greater scalability that are being considered. CUNY BPL is working to perpetuate and grow the use of BuildingCoach. As of this writing, about twenty buildings have either been connected or will be connected with operators coached / to be coached in the NYC municipal portfolio, twelve buildings across four campuses in NY State will use BuildingCoach, a NY upstate county wishes for six or seven buildings to participate with the support of funding from NYSERDA, and others have also expressed interest. In the decades to come, there will be an increasing percentage of large and mid-sized buildings that incorporate automated system optimization (ASO), and the building operators’ role will shift to spend more time on maintenance and monitoring. Meanwhile, programs such as BuildingCoach will play a critical role in optimizing operations. And, regardless of the emergence of ASO, operators will still need to understand how their systems operate so that they can monitor them properly. Within that context, BuildingCoach is an important step towards operators’ understanding of efficient building system operations.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Codiscovering graphical structure and functional relationships within data: A Gaussian Process framework for connecting the dots

Most problems within and beyond the scientific domain can be framed into one of the following three levels of complexity of function approximation. Type 1: Approximate an unknown function given input/output data. Type 2: Consider a collection of variables and functions, some of which are unknown, indexed by the nodes and hyperedges of a hypergraph (a generalized graph where edges can connect more than two vertices). Given partial observations of the variables of the hypergraph (satisfying the functional dependencies imposed by its structure), approximate all the unobserved variables and unknown functions. Type 3: Expanding on Type 2, if the hypergraph structure itself is unknown, use partial observations of the variables of the hypergraph to discover its structure and approximate its unknown functions. These hypergraphs offer a natural platform for organizing, communicating, and processing computational knowledge. While most scientific problems can be framed as the data-driven discovery of unknown functions in a computational hypergraph whose structure is known (Type 2), many require the data-driven discovery of the structure (connectivity) of the hypergraph itself (Type 3). We introduce an interpretable Gaussian Process (GP) framework for such (Type 3) problems that does not require randomization of the data, access to or control over its sampling, or sparsity of the unknown functions in a known or learned basis. Its polynomial complexity, which contrasts sharply with the super-exponential complexity of causal inference methods, is enabled by the nonlinear ANOVA capabilities of GPs used as a sensing mechanism.

Science & Technology - Other Topics↗

Large Language Model for Validation, Optical Calibration, and Learning (VOCAL) Distributed Temperature Sensing Interface

Distributed temperature sensing (DTS) using fiber optic sensors (FOS) offers a promising method for temperature measurements in advanced reactors, such as sodium fast reactors and molten salt cooled reactors. To support the calibration and validation of DTS measurements, Argonne National Laboratory developed the Validation, Optical Calibration, and Learning (VOCAL) software package. This report describes the integration of a local large language model (LLM) with a retrieval-augmented generation (RAG) system into the VOCAL interface to serve as an interactive user assistant. The LLM framework enhances the VOCAL platform’s accessibility to users by explaining interface components, clarifying inputs and outputs, and answering user queries dynamically in real-time. The accuracy of the LLM assistant performance was evaluated with 20 queries regarding the interface and its parameters using experimental data from the Thermal Hydraulic Experimental Test Article (THETA) facility. Results demonstrate that the LLM achieved a 95% accuracy rate, with a BERTScore of 0.8816 and SBERT value of 0.7417. Furthermore, validation of the RAG system within the LLM framework showed optimal accuracy with k-values between 1 and 2 using the k-refinement convergence test. The prompt perturbation analysis demonstrated good initial consistency for the RAG system, exhibiting the highest accuracy under punctuation variations and the greatest sensitivity under query reordering. Notably, the model’s errors were limited to data retrieval failures rather than factual hallucinations, reinforcing its baseline reliability. The integration of LLM provides a highly accurate, userfriendly enhancement to the VOCAL platform without disrupting its core computational capabilities for FOS calibration and validation.

Hong, Evan↗

Leveraging Cloud Platforms for Grid Modernization

Presentation held on Friday December 5th, 2025 at the “San Diego Tech Conference and Expo” about “Advanced Sensor Data Analytics and Cloud Computation for Grid Modernization”

24 POWER TRANSMISSION AND DISTRIBUTION↗

WM26 Paper Multi-Robot Collaboration for Hazardous Environments

Hazardous nuclear and industrial facilities are rarely designed for robots. Work in these domains demand precise manipulation and robust mobility in cluttered, constrained spaces where off-the-shelf platforms struggle and “one-size-fits-all” machines become costly and complex. Idaho National Laboratory (INL) is developing an autonomous, multi-robot inspection system that coordinates task-specific platforms rather than relying on a single omni-tool robot. An electric truck serves as a power and compute hub for a custom manipulator co-developed with Florida International University (FIU), a commercial mini crawler, a pan–tilt–zoom camera, and a Nexxis Argus LiDAR mapping system. Working in concert, these robots generate spatial, radiation, and temperature maps of the pit environments at the Hanford Waste Tank Farms. These systems will capture visual records and environmental telemetry to allow for analysis post inspection. The system architecture uses Robot Operating System 2 (ROS 2) for publish/subscribe integration, NVIDIA Isaac Sim and Unity for simulation and visualization, and algorithms such as NVBlox to fuse data into unified 3D overlays. This robot-agnostic approach reduces operator burden by enabling autonomy across heterogeneous platforms and lets each robot be used where it is strongest. Having autonomous functions means operators don’t have to fully control multiple different components. The ease of use could allow for more widespread adoption of advanced robotics at waste management sites that see continued use. By coordinating simpler, purpose-built mechanisms, the approach lowers design and manufacturing complexity, reduces capital risk in contaminated settings, and improves controllability for complex inspection and manipulation tasks. We present the architecture, early results, and lessons learned from building and deploying this coordinated multi-robot system, with the goal of accelerating safe, cost-effective adoption of advanced robotics at waste-management sites.

42 - ENGINEERING↗

spammR: an R package designed for analysis and integration of spatial multi-omic measurements

Spatial omics is a young and evolving field and as such shows rapid development of novel technologies and analysis methods to measure transcripts, proteins, metabolites, and post-translational modifications at high spatial resolution. These advances in technology have enabled the simultaneous generation of abundance profiles for multiple different omics types and associated microscopy imaging data, as well as their analysis in a spatial context. However, most analytical tools are designed for spatial transcriptomics platforms and are challenging to use in other contexts such as mass spectrometry-based measurements or metagenomics. To this end we present spammR (spatial analysis of multi-omics measurements in R), an R package that enables end-to-end analysis with a specific focus on mass-spectrometry derived spatial omics datasets with (1) smaller sample sizes and spatial sparsity of samples, (2) considerable missingness, and (3) no a-priori knowledge about proteins or genes of interest, relying on a fully data-driven approach.

spammR↗

Achieving Multimodal and Multicolor Luminescence in LaAlO 3 :Pr 3+ , Gd 3+ via Trap Engineering and Energy Transfer

Achieving multimodal luminescence within a single phosphor is vital for multifunctional applications but remains challenging due to complex color tuning and trap engineering. In this study, we report Pr 3+ and Gd 3+ co‐doped LaAlO 3 (LAO:PG) phosphors, designed through careful modulation of multilevel traps and Pr 3+ → Gd 3+ energy transfer dynamics. These materials exhibit diverse luminescence modes, including down‐conversion luminescence (DCL), up‐conversion luminescence (UCL), persistent luminescence (PersL), optically stimulated luminescence (OSL), and thermally stimulated luminescence (TSL) across a wide spectral range. Unlike previously studied Pr 3+ ‐doped LAO, the co‐doped LAO:PG shows DCL in both UV‐visible and NIR regions and displays ultraviolet‐C UCL under visible excitation. Notably, we observe, for the first time, PersL lasting several minutes in these phosphors—an improvement over the non‐PersL behavior of Pr 3+ ‐only doped LAO. Additionally, the LAO:PG phosphors exhibit strong OSL response. TSL analysis reveals five distinct trap levels linked to these properties. Density functional theory calculations further correlate intrinsic defects to these traps, supporting a proposed mechanism for the observed multimodal luminescence. These findings highlight LAO:PG as a promising platform for developing advanced phosphors with integrated luminescence modes, paving the way for future applications in data storage, phototherapy, and anti‐counterfeiting technologies.

Chemistry↗

Measurement of the 75 As ⁢(𝑛, 2⁢𝑛) cross section at 14.1 MeV at the National Ignition Facility

Here, a new measurement of the 75 As (𝑛, 2⁢𝑛) measurement was performed using the 14.1 MeV neutron pulse produced by fusion experiments at the National Ignition Facility (NIF). The target material consisted of GaAs foils encapsulated in aluminum irradiation containers, along with Au monitor foils, and positioned in the NIF chamber during high neutron yield shots. After irradiation, the GaAs and Au foils were counted with high purity germanium detectors to assess the production of (𝑛, 2⁢𝑛) activation products to determine the flux and the cross section of interest. The measured cross section was 966 ± 71 mb at 14.1 ± 0.37 MeV. This work provides a proof of concept for a platform for performing “ride-along” cross section measurements at NIF to contribute to existing nuclear data as well as measure new cross sections in the future.

and nuclear chemistry↗

Carbon-13 NMR spectra of lignin isolated from field grown transgenic poplar

Here we present a curated dataset of a series of 13C nuclear magnetic resonance (NMR) spectra of lignin isolated from transgenic monolignol 4-O-methyltransferase (MOMT4) engineered poplar. The transgenic poplar was collected from a 3-year field trial experiment. The poplar was Soxhlet-extracted with toluene/ethanol and the extractives-free poplar was then ball-milled in a Retsch PM100 planetary ball mill using a porcelain jar with ceramic balls at 600 rpm for 2 h. The ball-milled materials were then subjected to enzymatic hydrolysis for 48 h followed by centrifugation and washing with deionized water. The solid residue was extracted twice with 96:4 (v/v) 1,4-dioxane/water mixture at room temperature overnight. The extracts were combined, rotary evaporated, and freeze-dried to recover the lignin. The dry lignin samples were dissolved in deuterated dimethyl sulfoxide for NMR characterization. 13C experiments were performed in a Bruker Avance III HD 500 MHz NMR spectrometer operating at a frequency of 125.12 MHz for the 13C nucleus using a standard Bruker pulse sequence (zgpg) on a Prodigy platform cryoprobe. The NMR spectra were acquired under the following conditions: spectra width 229 ppm, 64k data points, 1s pulse delay, and 6k scans. All the data was processed using the Bruker’s TopSpin 3.6 software. Additional meta data is embedded in the raw spectra files.

13C NMR, lignin, poplar, field trial, MOMT4, CBI↗

CONTROL AND DATA ACQUISITION IN A CYBER-PHYSICAL MIDSTREAM TESTBED

This thesis presents the development of a laboratory-scale cyber–physical midstream pipeline testbed designed to address this gap and support research in industrial control systems security. The platform integrates pumps, valves, sensors, programmable logic controllers (PLCs), and a human–machine interface (HMI) to emulate the monitoring and control architecture of real pipeline operations. The physical process is implemented as a closed-loop liquid circulation system designed to replicate flow behavior characteristic of midstream pipeline infrastructure. The testbed enables real-time data acquisition of key process variables, including flow rate and pressure facilitating the generation of datasets representative of normal pipeline operation. A threat model encompassing common ICS attack vectors was developed, including sensor spoofing, command injection, false data injection, denial-of-service attacks, and relay manipulation. Multiple attack scenarios were implemented and evaluated to demonstrate how cyber intrusions targeting sensors, actuators, networks, and software propagate into measurable physical consequences in pipeline flow and pressure. The developed platform serves as a practical, cost-effective environment for experimentation, education, and future cybersecurity research in midstream pipeline systems.

42 ENGINEERING↗

Developing multi-gene CRISPRa/i programs to accelerate DBTL cycles in ABF hosts engineered for chemical production

This project developed and implemented a modular CRISPR activation and interference (CRISPRa/i) platform to accelerate strain optimization and pathway development for industrially relevant microbial hosts. By integrating multiplexed transcriptional perturbation tools with data-driven Design–Build–Test–Learn (DBTL) workflows, the team achieved reductions in cycle time and enhanced production of industrial aromatics, particularly 4-aminocinnamic acid (4-ACA), in Pseudomonas putida. Key accomplishments included: ● Development of a robust, tunable CRISPRa/i system in P. putida that enabled efficient multi-target gene regulation via guide RNA (gRNA) programs ● Completion of two full DBTL cycles, guided by machine learning (ML) models trained on transcriptomic and performance data, reducing engineering time by over 30% ● Optimization of multi-gene regulatory programs to balance expression of host and pathway modules, improve 4-ACA titers, and resolve metabolic bottlenecks ● Demonstration of system portability through a limited proof-of-concept extension in Acinetobacter baylyi, underscoring the generalizability of the approach ● Evaluation of strain performance on lignocellulosic biomass-derived substrates, demonstrating the feasibility of converting renewable carbon into aromatic building blocks These results illustrate the feasibility of applying ML-guided CRISPRa/i perturbation strategies to accelerate strain development in complex microbial systems. The resulting tools and datasets contribute to DOE objectives by improving platform predictability, reducing development costs, and enabling broader access to sustainable, economically viable bioproduction technologies.

09 BIOMASS FUELS↗