Search NASA⌕ Search

SEARCH · Search NASA

Results for “Genomic Engineering”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 199 records · Page 11

Engineering Escherichia coli for Urease-Driven Synthesis of Metal Oxide Nanomaterials

The development of functional nanomaterials with controlled morphologies is essential for advancements in medicine, electronics and computing, energy, catalysis, and environmental applications. However, conventional synthesis methods often demand high energy input and pose significant environmental challenges. Urease-based biomineralization presents an efficient, eco-friendly alternative for nanomaterial production under mild conditions. In this study, we engineered Escherichia coli ( E. coli ) to express a urease gene cluster from Sporosarcina pasteurii using CRAGE-Duet technology. The engineered strain successfully synthesized calcium carbonate and calcium phosphate crystals. Expanding the approach, we synthesized metal oxide nanoparticles, including hematite (Fe 2 O 3 ), and nanocrystalline anatase titanium dioxide (TiO 2 ). These nanomaterials were characterized by electron microscopy, demonstrating the potential of E. coli as a sustainable and versatile platform for green nanomaterial synthesis.

bacteria↗

Crop models: integrating systems from the molecular to global for agricultural productivity and sustainability

Mathematical models that simulate crop growth in response to environmental conditions and management practices are essential tools for exploring agriculture-based strategies to address food security and environmental sustainability challenges. Early applications of crop models focused on supporting farmers in making management decisions. Applications have since expanded to estimating future impacts on local and global food production from changing climates. Emerging applications of crop models aim to leverage how these models integrate plant processes across biological scales to identify engineering or breeding strategies that account for environmentally-responsive dynamics at field scales and for exploring solutions to improve sustainability. In this review, we highlight recent studies across these four broad application areas and highlight potential future directions for the crop modeling field.

Piao, Ximin [Univ. of Illinois at Urbana-Champaign↗

Impacts of Legacy and Contemporary Nitrogen Inputs on N 2 O and CO 2 Emissions in Miscanthus and Maize Cultivated Soils

ABSTRACT Nutrient inputs influence the sustainability of bioenergy crop production through contemporary (shortly after addition) and legacy effects (persisting over years) on microbial nitrogen (N) and carbon cycling, which contribute to greenhouse gas emissions. However, the relative importance of contemporary and legacy effects and how that could vary by crop functional types is poorly understood. Considering its rhizomatous roots and perennial growth, we hypothesized that Miscanthus × giganteus (M×g) would be more sensitive to legacy N fertilization and the historical context of its environment than an annual crop like maize. To test this hypothesis, we examined the effects of legacy and contemporary N inputs on nitrous oxide (N 2 O) and carbon dioxide (CO 2 ) emissions, as well as key N cycling genes in soils where M×g and maize were grown. A 150‐day soil incubation experiment was conducted using soils from a long‐term M×g and maize fertility experiment with three historic N fertilization rates (0, 112, and 336 kg N ha −1 year −1 ) and a contemporary amendment (60 mg N kg −1 ) with negative control (0 mg N kg −1 ). We observed significant increases in cumulative N 2 O emissions in Mxg soils relative to maize soils, particularly at higher legacy fertilization rates, while contemporary N had no significant effect. Bacterial amo A gene abundance, which plays a significant role in nitrification in nutrient‐rich soils, also increased with higher legacy fertilization rates in M×g soils but was unaffected by the contemporary N. In maize soils, legacy and contemporary N did not significantly affect N 2 O emissions, but cumulative CO 2 emissions and amo A gene abundance significantly increased. The abundances of nor B genes were not significantly influenced by either legacy fertilization or contemporary N amendments in either soil. Our findings demonstrate the greater importance of fertilization history over contemporary N in mediating soil N 2 O emissions, particularly for perennial bioenergy crops.

09 BIOMASS FUELS↗

Artificial intelligence tools for enzyme engineering and metabolic engineering

Enzyme engineering and metabolic engineering drive innovation in energy biotechnology. In recent years, artificial intelligence (AI) has supported successful applications in designing effective enzymes and productive microbial cell factories. This review summarizes recent advances in enzyme redesign using protein language models, de novo enzyme design with generative models, and AI tools for engineering metabolism and related cellular phenotypes. Across these areas, AI models are shifting from single modality inputs to integrated representations of protein function, metabolic pathways, and cell states. We emphasize that unifying the diverse data representations across scales will be necessary for advancements in energy biotechnology.

Volk, Michael [Univ. of Illinois at Urbana-Champai↗

Multi-omic characterization of a soil microbial consortium reveals critical role of succinate and glutamate metabolism during calcium carbonate precipitation

Microbially induced calcium carbonate precipitation (MICP) holds potential for use in soil stabilization and carbon sequestration, with the overall efficiency of the process being a major determinant for use in many environmental and civil engineering applications. While the biogeochemical pathways and enzymes driving MICP are known, the microbial metabolic networks and community dynamics underlying such precipitation remain poorly characterized. To address this gap, we developed a four-member consortium of soil bacteria (Curtobacterium flaccumfaciens, Rhodococcus qingshengii, Microbacterium sp., and Bacillus toyonensis), termed carbon storing consortium - A (CSC-A), that is capable of MICP. Prior work shows that MICP production is higher in CSC-A compared to the sum of carbonate produced by each member, suggesting carbonate production is driven by consortium dynamics. To that end we used a multi-omic integration approach of genomics, transcriptomics, and metabolomics to investigate potential inter-species interactions that may influence the MICP phenotype. Genomic life history characterizations identified evidence of niche specialization by B. toyonensis and Microbacterium, while metatranscriptomic analysis suggests R. qingshengii is a keystone species during growth in urea. By comparing individual species’ metabolomes to the metabolic profile of a shared well of precipitated metabolites, we identified over 200 metabolites predicted to be produced or consumed by CSC-A members. Integrating both data types to search the KEGG reactome highlighted a network centered around glutamine metabolism and branched chain amino acid biosynthesis under regulation during CSC-A growth in urea. Succinate metabolism was also a major node in this network and laboratory assays confirmed that increasing the amount of succinate in the growth medium leads to increased carbonate precipitation by CSC-A, a critical confirmation of our modeling approach. By isolating and identifying the interconnected metabolic components underlying MICP in CSC-A, we identified keystone taxa, metabolites, and pathways important for future optimization of the application of this consortia to carbonate precipitation.

carbon storing consortium - A (CSC-A)↗

Data for "Land-based Resources for Engineered Carbon Dioxide Removal in the United States Exceed the Expected Needs"

Gigatonne-scale atmospheric carbon dioxide removal (CDR), alongside deep emission cuts, is critical to stabilizing the climate. However, some of the most scalable CDR technologies are also the most land intensive. Here, we examine whether adequate land resources exist in the contiguous United States to meet CDR targets when prioritizing grid emissions reduction, food production, and the protection of sensitive ecosystems. We focus on biomass carbon removal and storage (BiCRS) and direct air capture and storage (DACS) and show that suitable lands exceed the expected needs: 37.6 million hectares of land are available for BiCRS, resulting in 0.26 GtCO2 of CDR/year, and 34 million hectares are suitable for wind- and solar-powered DACS, resulting in 4.8 GtCO2 of CDR/year if facilities are co-located with geologic CO2 storage. We identify biomass and energy supply hotspots to meet CDR targets while ensuring land protection and minimizing land competition.

carbon↗

Evaluating the potential of disaggregated memory systems for HPC applications

Summary Disaggregated memory is a promising approach that addresses the limitations of traditional memory architectures by enabling memory to be decoupled from compute nodes and shared across a data center. Cloud platforms have deployed such systems to improve overall system memory utilization, but performance can vary across workloads. High‐performance computing (HPC) is crucial in scientific and engineering applications, where HPC machines also face the issue of underutilized memory. As a result, improving system memory utilization while understanding workload performance is essential for HPC operators. Therefore, learning the potential of a disaggregated memory system before deployment is a critical step. This paper proposes a methodology for exploring the design space of a disaggregated memory system. It incorporates key metrics that affect performance on disaggregated memory systems: memory capacity, local and remote memory access ratio, injection bandwidth, and bisection bandwidth, providing an intuitive approach to guide machine configurations based on technology trends and workload characteristics. We apply our methodology to analyze thirteen diverse workloads, including AI training, data analysis, genomics, protein, fusion, atomic nuclei, and traditional HPC bookends. Our methodology demonstrates the ability to comprehend the potential and pitfalls of a disaggregated memory system and provides motivation for machine configurations. Our results show that eleven of our thirteen applications can leverage injection bandwidth disaggregated memory without affecting performance, while one pays a rack bisection bandwidth penalty and two pay the system‐wide bisection bandwidth penalty. In addition, we also show that intra‐rack memory disaggregation would meet the application's memory requirement and provide enough remote memory bandwidth.

Ding, Nan↗

Activity-targeted metaproteomics uncovers rare syntrophic bacteria central to anaerobic community metabolism

Syntrophic microbial consortia can contribute significantly to the activity and function of anoxic ecosystems, yet are often too rare to study their in situ physiologies using traditional molecular methods. Here, in this study, we describe a technical innovation combining bioorthogonal non-canonical amino acid tagging (BONCAT), stable isotope probing, and metaproteomics to improve the recovery of proteins from active community members and track isotope incorporation. Both click chemistry-enabled cell-sorting and direct protein pulldown coupled to metaproteomics improved recovery of isotopically labeled proteins during acetate oxidation within a full-scale anaerobic digester. Resulting labeled protein expression profiles revealed elevated activity of a rare and uncharacterized syntrophic bacterium belonging to the family Natronincolaceae. BONCAT-based capture of newly translated proteins provided direct molecular evidence for the expression of a previously hypothesized oxidative glycine pathway for syntrophic acetate oxidation by this microorganism, showcasing the potential of targeted metaproteomics to characterize rare and active cells central to community metabolism in natural and engineered ecosystems.

Friedline, Skyler [Univ. of British Columbia, Vanc↗

Specialization Restricts the Evolutionary Paths Available to Yeast Sugar Transporters

Functional innovation at the protein level is a key source of evolutionary novelties. The constraints on functional innovations are likely to be highly specific in different proteins, which are shaped by their unique histories and the extent of global epistasis that arises from their structures and biochemistries. These contextual nuances in the sequence–function relationship have implications both for a basic understanding of the evolutionary process and for engineering proteins with desirable properties. Here, we have investigated the molecular basis of novel function in a model member of an ancient, conserved, and biotechnologically relevant protein family. These Major Facilitator Superfamily sugar porters are a functionally diverse group of proteins that are thought to be highly plastic and evolvable. By dissecting a recent evolutionary innovation in an α-glucoside transporter from the yeast Saccharomyces eubayanus, we show that the ability to transport a novel substrate requires high-order interactions between many protein regions and numerous specific residues proximal to the transport channel. To reconcile the functional diversity of this family with the constrained evolution of this model protein, we generated new, state-of-the-art genome annotations for 332 Saccharomycotina yeast species spanning ~400 My of evolution. By integrating phylogenetic and phenotypic analyses across these species, we show that the model yeast α-glucoside transporters likely evolved from a multifunctional ancestor and became subfunctionalized. The accumulation of additive and epistatic substitutions likely entrenched this subfunction, which made the simultaneous acquisition of multiple interacting substitutions the only reasonably accessible path to novelty.

59 BASIC BIOLOGICAL SCIENCES↗

Evaluation of two inoculation routes of an adenovirus-mediated viral protein inhibitor in a Crimean-Congo hemorrhagic fever mouse model

Crimean-Congo hemorrhagic fever virus (CCHFV) is a tick-borne nairovirus with a wide geographic spread that can cause severe and lethal disease. No specific medical countermeasures are approved to combat this illness. The CCHFV L protein contains an ovarian tumor (OTU) domain with a cysteine protease thought to modulate cellular immune responses by removing ubiquitin and ISG15 post-translational modifications from host and viral proteins. Viral deubiquitinases like CCHFV OTU are attractive drug targets, as blocking their activity may enhance cellular immune responses to infection, and potentially inhibit viral replication itself. We previously demonstrated that the engineered ubiquitin variant CC4 is a potent inhibitor of CCHFV replication in vitro. A major challenge of the therapeutic use of small protein inhibitors such as CC4 is their requirement for intracellular delivery, e.g., by viral vectors. In this study, we examined the feasibility of in vivo CC4 delivery by a replication-deficient recombinant adenovirus (Ad-CC4) in a lethal CCHFV mouse model. Since the liver is a primary target of CCHFV infection, we aimed to optimize delivery to this organ by comparing intravenous (tail vein) and intraperitoneal injection of Ad-CC4. While tail vein injection is a traditional route for adenovirus delivery, in our hands intraperitoneal injection resulted in higher and more widespread levels of adenovirus genome in tissues, including, as intended, the liver. However, despite promising in vitro results, neither route of in vivo CC4 treatment resulted in protection from a lethal CCHFV infection.

59 BASIC BIOLOGICAL SCIENCES↗

Synthetic Biology PacBio/JAWS QC Analysis (PBJ) v3.0

This software was designed as a sequence validation tool for the assembly of synthetic constructs. It analyzes FASTQ files against a list of reference sequences, combining the results from eight sequencing libraries to generate a summary, and the files needed to view the results in the Integrative Genomics Viewer (IGV) application for manual verification. This was developed for FASTQ files generated by PacBio sequencing, but could be used on any FASTQ files that do not have paired end reads. It can be used to analyze one - eight libraries at a time, and assumes that each construct sequence in the reference will be in each pool, however, this is not a requirement. This is used to identify which libraries of pooled sequences contains a perfect match, or fixable match to the reference file. This pipeline uses many freely available open source libraries, the value added is that in our application the steps of the pipeline are defined in Workflow Description Language (WDL) and run through the Cromwell workflow engine in Docker containers, for easy distribution and set up, as well as the user friendly html summary that is generated.

Simirenko, Lisa↗

Identification and overexpression of endogenous transcription factors to enhance lipid accumulation in the biotechnologically relevant species Chlamydomonas pacifica

Sustainable low-carbon energy solutions are critical to mitigating global carbon emissions. Algae-based platforms offer potential by converting carbon dioxide into valuable products while aiding carbon sequestration. However, scaling algae cultivation faces challenges like contamination in outdoor systems. Previously, our lab evolved Chlamydomonas pacifica, an extremophile green alga, which tolerates high temperature, pH, salinity, and light, making it ideal for large-scale bioproduct production, including biodiesel. Here, we enhanced lipid accumulation in evolved C. pacifica by identifying and overexpressing key endogenous transcription factors through genome-wide in-silico analysis and in-vivo testing. These factors include Lipid Remodeling Regulator 1 (CpaLRL1), Nitrogen Response Regulator 1 (CpaNRR1), Compromised Hydrolysis of Triacylglycerols 7 (CpaCHT7), and Phosphorus Starvation Response 1 (CpaPSR1). Under nitrogen deprivation, CpaLRL1, CpaNRR1, and CpaCHT7 overexpression enhanced lipid accumulation compared to wild-type. However, CpaPSR1 increased lipid accumulation compared to wild-type in normal media and did not increase further under nitrogen deprivation, highlighting the difference in function based on media conditions. Notably, lipid analysis of CpaPSR1 under normal media conditions revealed a 2.4-fold increase in triglycerides (TAGs) compared to the wild-type, highlighting its potential for biodiesel production. This approach provides a framework for transcription factor-focused metabolic engineering in algae, advancing bioenergy and biomaterial production.

Biofuels↗

A substrate-multiplexed platform for profiling enzymatic potential of plant family 1 glycosyltransferases

Plants have expanded various biosynthetic enzyme families to produce a wide diversity of natural products; however, most enzymes encoded in plant genomes remain uncharacterized, highlighting the need for new functional genomic approaches. Here, we report a platform enabling the rapid functional characterization of plant family 1 glycosyltransferases, which serve important roles in plant development, defense, and communication. Using substrate-multiplexed reactions, mass spectrometry, and automated analysis, we screen 85 enzymes against a diverse library of 453 natural products, for a total of nearly 40,000 possible reactions. The resulting dataset reveals a widespread promiscuity and a strong preference for planar, hydroxylated aromatic substrates among family 1 glycosyltransferases. We also characterize glycosyltransferases with an unusually wide substrate scope and with a non-canonical Cys-Asp catalytic dyad. This work establishes a widely-applicable enzymatic screening pipeline, reflects the immense glycosylation capability of plants, and has implications in biocatalysis, metabolic engineering, and gene discovery.

Sirirungruang, Sasilada↗

SynBio QC Dual Barcode QC (DBC) v1.0

This software was designed as a sequence validation tool for the assembly of synthetic constructs, where the constructs have a high degree of similarity and thus are barcoded prior to the sequencing library prep. It demultiplexes each FASTQ file for each barcode, then analyzes the resulting FASTQ files against a list of reference sequences for that barcode/library, combining the results from eight sequencing libraries to generate a summary, and the files needed to view the results in the Integrative Genomics Viewer (IGV) application for manual verification. This was developed for FASTQ files generated by PacBio sequencing, but could be used on any FASTQ files that do not have paired end reads. It can be used to analyze one - eight libraries at a time. Each construct is independently analyzed with only the sequences with the same barcode, in the same pooled library. Then the results are combined into a user friendly summary. This is used to identify which libraries of pooled sequences contains a perfect match, or fixable match to the reference file. This pipeline uses many freely available open source libraries, the value added is that in our application the steps of the pipeline are defined in Workflow Description Language (WDL) and run through the Cromwell workflow engine in Docker containers, for easy distribution and set up, as well as the user friendly html summary that is generated.

Simirenko, Lisa↗

Multi‐Omics Analyses Reveal Divergent Molecular Mechanisms Underlying Plant Biomass Conversion by Five Fungi

Fungal plant biomass conversion (FPBC) is of great importance to the global carbon cycle and has been increasingly applied for the production of biofuel and biochemicals from lignocellulose. However, the comprehensive understanding of relevant molecular mechanisms in different fungi remains challenging. Here, we comparatively analyzed the transcriptome, proteome and metabolome profile of four ascomycetes and one basidiomycete fungi during their growth on two common agricultural feedstocks (soybean hulls and corn stover). We revealed strong time‐, substrate‐ and species‐specific responses at multi‐omics levels for the tested fungi, highlighting species‐specific carbon utilization approaches and evolutionary adaptation to environmental niches. Notably, a remarkable expressional diversity of lignocellulose degrading enzymes, sugar transporter and metabolic genes, as well as industrially relevant metabolites were identified across different fungi and cultivation conditions. The findings improves our understanding of complex molecular networks underlying FPBC and fungal ecological roles, offering novel insights that can guide future genetic engineering of fungi for valorization of agriculture waste into value‐added bioproducts.

CAZy↗

Metagenomes and metagenome-assembled genomes from microbial communities in a biological nutrient removal plant operated at Hamptons Road Sanitation District (HRSD) with high and low dissolved oxygen conditions

Aeration is a major cost at biological nutrient removal (BNR) plants. We report on microbial communities in a pilot-scale BNR system before and after a dissolved oxygen transition from 2.5 to 0.2 mg/L implemented over 18 months. Four PacBio metagenomes and 316 metagenome-assembled genomes are announced.

dissolved oxygen↗

A multi-omic characterization of the physiological responses to salt stress in Scenedesmus obliquus UTEX393

Scenedesmus obliquus UTEX393 is a promising microalgal candidate for sustainable biomanufacturing but its limited halotolerance hinders large-scale cultivation in saline environments. To investigate the molecular basis of salt stress responses, we conducted a comprehensive multi-omic analysis integrating genomics, transcriptomics, proteomics, lipidomics, metabolomics, and DNA affinity purification sequencing (DAP-seq). An improved nuclear genome assembly and annotation yielded 19,017 gene models and a 97% BUSCO completeness score, enabling construction of a genome-scale metabolic model. Comparing 15 ppt salinity stress to 5 ppt control, growth and productivity were significantly reduced, accompanied by widespread transcriptomic and proteomic changes. Transcriptomic analysis revealed downregulation of photosynthetic machinery and energy conservation genes, and upregulation of stress-responsive elements such as expansins, flavodoxins, and osmoprotectants. Lipidomic profiling showed accumulation of triacylglycerols (TAGs) and degradation of galactosyl lipids, consistent with a shift toward lipid biosynthesis to mitigate redox imbalance. Depletion of key polar metabolites and branched-chain amino acids suggested a rerouting of central carbon metabolism under stress. DAP-seq identified key transcription factors, including LHY1 and SPL12, that target central metabolic enzymes involved in redox balancing, such as glyceraldehyde-3-phosphate dehydrogenase (GAPDH) and malate dehydrogenase (MDH). These findings establish a regulatory-metabolic framework linking redox stress to lipid accumulation and reveal potential engineering targets to enhance salt tolerance. Overall, the multi-omic analysis supports the “overflow” hypothesis, where impaired photosynthesis results in excess reducing equivalents being diverted into TAG synthesis and highlights transcriptional regulators as candidates for improving algal robustness in brackish environments.

09 BIOMASS FUELS↗

Tetranucleotide frequencies differentiate genomic boundaries and metabolic strategies across environmental microbiomes

Microbiomes are constrained by physicochemical conditions, nutrient regimes, and community interactions across diverse environments, yet genomic signatures of this adaptation remain unclear. Metagenome sequencing is a powerful technique to analyze genomic content in the context of natural environments, establishing concepts of microbial ecological trends. Here, we developed a data discovery tool-a tetranucleotide-informed metagenome stability diagram-that is publicly available in the integrated microbial genomes and microbiomes (IMG/M) platform for metagenome ecosystem analyses. We analyzed the tetranucleotide frequencies from quality-filtered and unassembled sequence data of over 12,000 metagenomes to assess ecosystem-specific microbial community composition and function. We found that tetranucleotide frequencies can differentiate communities across various natural environments and that specific functional and metabolic trends can be observed in this structuring. Our tool places metagenomes sampled from diverse environments into clusters and along gradients of tetranucleotide frequency similarity, suggesting microbiome community compositions specific to gradient conditions. Within the resulting metagenome clusters, we identify protein-coding gene identifiers that are most differentiated between ecosystem classifications. We plan for annual updates to the metagenome stability diagram in IMG/M with new data, allowing for refinement of the ecosystem classifications delineated here. This framework has the potential to inform future studies on microbiome engineering, bioremediation, and the prediction of microbial community responses to environmental change. IMPORTANCE: Microbes adapt to diverse environments influenced by factors like temperature, acidity, and nutrient availability. We developed a new tool to analyze and visualize the genetic makeup of over 12,000 microbial communities, revealing patterns linked to specific functions and metabolic processes. This tool groups similar microbial communities and identifies characteristic genes within environments. By continually updating this tool, we aim to advance our understanding of microbial ecology, enabling applications like microbial engineering, bioremediation, and predicting responses to environmental change.

Kellom, Matthew↗