Search NASA⌕ Search

SEARCH · Search NASA

Results for “Genomic Engineering”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 181 records · Page 10

Beyond Component Optimization: Systems Level Biodesign for Lanthanide Recovery

Global demand for lanthanides (Ln) is projected to rise sharply over the next decade, while geographically concentrated supply chains and the low concentrations and matrix complexity of secondary feedstocks limit the reach of conventional hydro- and pyrometallurgical separation. Engineered biological systems offer a selective, low-energy alternative, and component-level advances in Ln-binding proteins, AI-designed selective scaffolds, and cell-surface display platforms now rival synthetic chelators in affinity and selectivity. These components, however, remain functionally isolated. Currently, there are no engineered chassis coupling recognition, intracellular trafficking, accumulation, and controlled release into an end-to-end pipeline. Here, we outline how new biodesign strategies and chassis selection must move beyond bioleaching to encompass the full recovery pathway. Achieving this requires integrating AI/ML-guided design, genome-scale build tools, high-throughput phenotyping, and biophysical transport modeling within a Design–Build–Test–Learn cycle tuned to recognition, trafficking, accumulation, and release.

Biodesign↗

Data for Rewiring Yeast Metabolism for Producing 2,3-Butanediol and Two Downstream Applications: Techno-Economic Analysis and Life Cycle Assessment of Methyl Ethyl Ketone (MEK) and Agricultural Biostimulant Production

Rising concerns for sustainability and global climate change have driven the development of sustainable production pathways for biofuels and chemicals from lignocellulosic biomass via integrated biological and chemical processes. We constructed an engineered Saccharomyces cerevisiae capable of producing 2,3-butanediol (2,3-BDO) from glucose without accumulating ethanol and glycerol, which hinder downstream processing of 2,3-BDO, through extensive metabolic reprogramming. Specifically, we introduced heterologous 2,3-BDO biosynthetic enzymes and deleted the major isozymes of ethanol and glycerol biosynthetic enzymes. In addition, we introduced an NAD+ regenerating Pyruvate-Malate (PM) cycle and enhanced the NAD+ regenerating capability of the PM cycle to resolve the redox imbalance from the deletion of ethanol and glycerol production pathways. The resulting engineered yeast produced 109.9 g/L of 2,3-BDO with a productivity of 1.0 g/L/h and a yield of 0.36 g/g glucose in a fed-batch fermentation. We also conducted techno-economic analysis (TEA) and life cycle assessment (LCA) of the production of methyl ethyl ketone (MEK) through catalytic dehydration of 2,3-BDO. A TEA based on the experimental results indicated that the minimum product selling price (MPSP) was estimated to be $1.90/kg. Regarding cradle-to-grave LCA, 100-year global warming potential (GWP100) and fossil energy consumption (FEC) were found to be 0.37 kg CO2 eq/kg and 3.1 MJ/kg, respectively. These results demonstrated the feasibility of cost-competitive and sustainable bio-based MEK production via yeast fermentation. In addition, we explored the possibility of using the fermentation broth containing 2,3-BDO as a biostimulant inducing drought tolerance in plants. As a result, the yeast 2,3-BDO fermentation broth can induce drought tolerance in Arabidopsis thaliana without a complicated purification process.

Economics↗

Omics-driven onboarding of the carotenoid producing red yeast Xanthophyllomyces dendrorhous CBS 6938

Transcriptomics is a powerful approach for functional genomics and systems biology, yet it can also be used for genetic part discovery. Here, we derive constitutive and light-regulated promoters directly from transcriptomics data of the basidiomycete red yeast Xanthophyllomyces dendrorhous CBS 6938 (anamorph Phaffia rhodozyma) and use these promoters with other genetic elements to create a modular synthetic biology parts collection for this organism. X. dendrorhous is currently the sole biotechnologically relevant yeast in the Tremellomycete class-it produces large amounts of astaxanthin, especially under oxidative stress and exposure to light. Thus, we performed transcriptomics on X. dendrorhous under different wavelengths of light (red, green, blue, and ultraviolet) and oxidative stress. Differential gene expression analysis (DGE) revealed that terpenoid biosynthesis was primarily upregulated by light through crtI, while oxidative stress upregulated several genes in the pathway. Further gene ontology (GO) analysis revealed a complex survival response to ultraviolet (UV) where X. dendrorhous upregulates aromatic amino acid and tetraterpenoid biosynthesis and downregulates central carbon metabolism and respiration. The DGE data was also used to identify 26 constitutive and regulated genes, and then, putative promoters for each of the 26 genes were derived from the genome. Simultaneously, a modular cloning system for X. dendrorhous was developed, including integration sites, terminators, selection markers, and reporters. Each of the 26 putative promoters were integrated into the genome and characterized by luciferase assay in the dark and under UV light. The putative constitutive promoters were constitutive in the synthetic genetic context, but so were many of the putative regulated promoters. Notably, one putative promoter, derived from a hypothetical gene, showed ninefold activation upon UV exposure. Thus, this study reveals metabolic pathway regulation and develops a genetic parts collection for X. dendrorhous from transcriptomic data. Therefore, this study demonstrates that combining systems biology and synthetic biology into an omics-to-parts workflow can simultaneously provide useful biological insight and genetic tools for nonconventional microbes, particularly those without a related model organism. This approach can enhance current efforts to engineer diverse microbes.

60 APPLIED LIFE SCIENCES↗

Stage-resolved gene regulatory network analysis reveals developmental reprogramming and genes with robust stem-preferred expression in sorghum

Sorghum bicolor is a deep-rooted, heat- and drought-tolerant crop that thrives on marginal lands and is increasingly valued for its applications in biofuel, bioenergy, and biopolymer production. The sorghum stem, which can reach 4–5 m in length, serves as the primary reservoir of both lignocellulosic biomass and soluble sugars, making it a promising bioenergy feedstock. Although recent advances in genetic, genomic, and transcriptomic resources have improved our understanding of sorghum biology, comprehensive genome-wide analyses of functional dynamics across diverse organ types and developmental stages remain limited. In particular, candidate genes with stem preferred expression pattern or their associated cis-regulatory elements, which may program key stem-related functions and enable organ- or tissue-specific engineering, have not yet been identified.

59 BASIC BIOLOGICAL SCIENCES↗

Discovery, characterization, and application of chromosomal integration sites in the hyperthermophilic archaeon Sulfolobus islandicus

Sulfolobus islandicus , an emerging archaeal model organism, offers unique advantages for metabolic engineering and synthetic biology applications owing to its ability to thrive in extreme environments. Although several genetic tools have been established for this organism, the lack of well-characterized chromosomal integration sites has limited its potential as a cellular factory. Here, in this work, we systematically identified and characterized 13 artificial CRISPR RNAs targeting eight integration sites in S. islandicus using the CRISPR-COPIES pipeline and a multi-omics-informed computational workflow. We leveraged the endogenous CRISPR-Cas system to integrate the reporter gene lacS and validated heterologous expression through a β-galactosidase assay, revealing significant positional effects. As a proof of concept, we utilized these sites to genetically manipulate lipid ether composition by overexpressing glycerol dibiphytanyl glycerol tetraether (GDGT) ring synthase B (GrsB). This study expands the genetic toolbox for S. islandicus and advances its potential as a robust platform for archaeal synthetic biology and industrial biotechnology.

59 BASIC BIOLOGICAL SCIENCES↗

Rapid and efficient in planta genome editing in sorghum using foxtail mosaic virus‐mediated sgRNA delivery

SUMMARY The requirement of in vitro tissue culture for the delivery of gene editing reagents limits the application of gene editing to commercially relevant varieties of many crop species. To overcome this bottleneck, plant RNA viruses have been deployed as versatile tools for in planta delivery of recombinant RNA. Viral delivery of single‐guide RNAs (sgRNAs) to transgenic plants that stably express CRISPR‐associated (Cas) endonuclease has been successfully used for targeted mutagenesis in several dicotyledonous and few monocotyledonous plants. Progress with this approach in monocotyledonous plants is limited so far by the availability of effective viral vectors. We engineered a set of foxtail mosaic virus (FoMV) and barley stripe mosaic virus (BSMV) vectors to deliver the fluorescent protein AmCyan to track viral infection and movement in Sorghum bicolor . We further used these viruses to deliver and express sgRNAs to Cas9 and Green Fluorescent Protein (GFP) expressing transgenic sorghum lines, targeting Phytoene desaturase ( PDS ), Magnesium‐chelatase subunit I ( MgCh ), 4‐hydroxy‐3‐methylbut‐2‐enyl diphosphate reductase , orthologs of maize Lemon white1 ( Lw1 ) or GFP . The recombinant BSMV did neither infect sorghum nor deliver or express AmCyan and sgRNAs. In contrast, the recombinant FoMV systemically spread throughout sorghum plants and induced somatic mutations with frequencies reaching up to 60%. This mutagenesis led to visible phenotypic changes, demonstrating the potential of FoMV for in planta gene editing and functional genomics studies in sorghum.

54 ENVIRONMENTAL SCIENCES↗

Engineering Microbial Communities: Frontier Science for the Bioeconomy Workshop Series

In nature, biological systems are shaped by complex interactions of diverse microorganisms such as bacteria, archaea, fungi, and viruses living within communities called microbiomes (Berg et al. 2020; Prescott 2017). These collective interactions result in emergent community properties that can be leveraged for beneficial purposes such as bioenergy and biomolecule production. Given this potential and the immensity of microbial genomic diversity, the U.S. Department of Energy’s (DOE) Biological and Environmental Research (BER) program has long invested in research to better understand the biology of environmental microbes and microbiomes.

59 BASIC BIOLOGICAL SCIENCES↗

High-throughput protein characterization by complementation using DNA barcoded fragment libraries

Abstract Our ability to predict, control, or design biological function is fundamentally limited by poorly annotated gene function. This can be particularly challenging in non-model systems. Accordingly, there is motivation for new high-throughput methods for accurate functional annotation. Here, we used co mplementation of aux otrophs and DNA barcode seq uencing (Coaux-Seq) to enable high-throughput characterization of protein function. Fragment libraries from eleven genetically diverse bacteria were tested in twenty different auxotrophic strains of Escherichia coli to identify genes that complement missing biochemical activity. We recovered 41% of expected hits, with effectiveness ranging per source genome, and observed success even with distant E. coli relatives like Bacillus subtilis and Bacteroides thetaiotaomicron . Coaux-Seq provided the first experimental validation for 53 proteins, of which 11 are less than 40% identical to an experimentally characterized protein. Among the unexpected function identified was a sulfate uptake transporter, an O-succinylhomoserine sulfhydrylase for methionine synthesis, and an aminotransferase. We also identified instances of cross-feeding wherein protein overexpression and nearby non-auxotrophic strains enabled growth. Altogether, Coaux-Seq’s utility is demonstrated, with future applications in ecology, health, and engineering.

59 BASIC BIOLOGICAL SCIENCES↗

Catabolism of β-5 linked aromatics by Novosphingobium aromaticivorans

ABSTRACT Aromatic compounds are an important source of commodity chemicals traditionally produced from fossil fuels. Aromatics derived from plant lignin can potentially be converted into commodity chemicals through depolymerization followed by microbial funneling of monomers and low molecular weight oligomers. This study investigates the catabolism of the β-5 linked aromatic dimer dehydrodiconiferyl alcohol (DC-A) by the bacterium Novosphingobium aromaticivorans . We used genome-wide screens to identify candidate genes involved in DC-A catabolism. Subsequent in vivo and in vitro analyses of these candidate genes elucidated a catabolic pathway composed of four required gene products and several partially redundant dehydrogenases that convert DC-A to aromatic monomers that can be funneled into the central aromatic metabolic pathway of N. aromaticivorans . Specifically, a newly identified γ-formaldehyde lyase, PcfL, opens the phenylcoumaran ring to form a stilbene and formaldehyde. A lignostilbene dioxygenase, LsdD, then cleaves the stilbene to generate the aromatic monomers vanillin and 5-formylferulate (5-FF). We also showed that the aldehyde dehydrogenase FerD oxidizes 5-FF before it is decarboxylated by LigW, yielding ferulic acid. We found that some enzymes involved in the β-5 catabolism pathway can act on multiple substrates and that some steps in the pathway can be mediated by multiple enzymes, providing new insights into the robust flexibility of aromatic catabolism in N. aromaticivorans . A comparative genomic analysis predicted that the newly discovered β-5 aromatic catabolic pathway is common within the order Sphingomonadales. IMPORTANCE In the transition to a circular bioeconomy, the plant polymer lignin holds promise as a renewable source of industrially important aromatic chemicals. However, since lignin contains aromatic subunits joined by various chemical linkages, producing single chemical products from this polymer can be challenging. One strategy to overcome this challenge is using microbes to funnel a mixture of lignin-derived aromatics into target chemical products. This approach requires strategies to cleave the major inter-unit linkages of lignin to release monomers for funneling into valuable products. In this study, we report newly discovered aspects of a pathway by which the Novosphingobium aromaticivorans DSM12444 catabolizes aromatics joined by the second most common inter-unit linkage in lignin, the β-5 linkage. This work advances our knowledge of aromatic catabolic pathways, laying the groundwork for future metabolic engineering of this and other microbes for optimized conversion of lignin into products.

59 BASIC BIOLOGICAL SCIENCES↗

Rewinding evolution in planta: A Rubisco-null platform validates high-performance ancestral enzymes

Improving the photosynthetic enzyme Rubisco is a key target for enhancing C3 crop productivity, but progress has been hampered by the difficulty of evaluating engineered variants in planta without interference from the native enzyme. Here, we report the creation of a Rubisco-null Nicotiana tabacum platform by using CRISPR-Cas9 to knock out all 11 nuclear-encoded small subunit (rbcS) genes. Knockout was achieved in a line expressing cyanobacterial Rubisco from the plastid genome, allowing the recovery of viable plants. We then developed a chloroplast expression system for coexpressing both large and small subunits from the plastid genome. We expressed two resurrected ancestral Rubiscos from the Solanaceae family. The resulting transgenic plants were phenotypically normal and accumulated Rubisco to wild-type levels. Importantly, kinetic analyses of the purified ancestral enzymes revealed they possessed a 16 to 20% higher catalytic efficiency (k cat,air /K c,air ) under ambient conditions, driven by a significantly faster turnover rate (k cat,air ). We have demonstrated that our system allows robust in vivo assessment of novel Rubiscos and that ancestral reconstruction is a powerful strategy for identifying superior enzymes to improve photosynthesis in C3 crops.

59 BASIC BIOLOGICAL SCIENCES↗

Biomolecular Analysis Capability for Cellular and Omics Research on the International Space Station

International Space Station (ISS) assembly complete ushered a new era focused on utilization of this state-of-the-art orbiting laboratory to advance science and technology research in a wide array of disciplines, with benefits to Earth and space exploration. ISS enabling capability for research in cellular and molecular biology includes equipment for in situ, on-orbit analysis of biomolecules. Applications of this growing capability range from biomedicine and biotechnology to the emerging field of Omics. For example, Biomolecule Sequencer is a space-based miniature DNA sequencer that provides nucleotide sequence data for entire samples, which may be used for purposes such as microorganism identification and astrobiology. It complements the use of WetLab-2 SmartCycler"TradeMark", which extracts RNA and provides real-time quantitative gene expression data analysis from biospecimens sampled or cultured onboard the ISS, for downlink to ground investigators, with applications ranging from clinical tissue evaluation to multigenerational assessment of organismal alterations. And the Genes in Space-1 investigation, aimed at examining epigenetic changes, employs polymerase chain reaction to detect immune system alterations. In addition, an increasing assortment of tools to visualize the subcellular distribution of tagged macromolecules is becoming available onboard the ISS. For instance, the NASA LMM (Light Microscopy Module) is a flexible light microscopy imaging facility that enables imaging of physical and biological microscopic phenomena in microgravity. Another light microscopy system modified for use in space to image life sciences payloads is initially used by the Heart Cells investigation ("Effects of Microgravity on Stem Cell-Derived Cardiomyocytes for Human Cardiovascular Disease Modeling and Drug Discovery"). Also, the JAXA Microscope system can perform remotely controllable light, phase-contrast, and fluorescent observations. And upcoming confocal microscopy capability will allow for optical sectioning of biological tissues to determine microanatomical localization of biomarkers. Furthermore, NASA's geneLAB effort addresses integration of genomic, epigenomic, transcriptomic, proteomic and metabolomic datasets, by applying an innovative open source science platform for multi-investigator high throughput utilization of the ISS. In sum, the expanding ISS capability for analysis of biomolecules is enabling innovative research in a broad spectrum of areas such as cellular and molecular biology, biotechnology, tissue engineering, biomedicine, and Omics, providing manifold benefits for humanity.

Guinart-Ramirez, Y.↗

Probing the limits of genetic recoding using multi-omics-guided evolution

Engineering the genetic code—by reassigning multiple of the 64 natural codons—enables making organisms resistant to all viruses, preventing genetic information exchange, and allowing the biosynthesis of genetically encoded unnatural polymers. However, synonymous codon replacement—recoding—is frequently lethal, and how recoding impacts fitness remains poorly explored. Here, we explore these effects using genome synthesis, directed evolution, and genome-transcriptome-translatome-proteome co-profiling on multiple synthetic Escherichia coli genomes. We construct six partially recoded E. coli strains bearing up to 45.8% of a synthetic genome with a deleterious 57-codon genetic code. As our analyses revealed widespread defects—including unassigned codons in Syn61 and Syn57—we apply multi-omics to revise our genome design and mitigate defects. Using multi-omics, we show that recoding induces transcriptional and translational changes leading to fitness defects under hundreds of conditions. Finally, we develop a multi-omics-guided evolution strategy that rapidly restores fitness, enabling genome synthesis with radical changes.

Nyerges, Akos [Harvard Medical School, Boston, MA ↗

EvoNet: A phylogenomic and systems biology approach to identify genes underlying plant survival in marginal, low‐N soils

The DOE‐BER “EvoNet” project investigates the genetic and molecular basis of plant resilience in extreme environments. We do this by identifying key genes that enable “extreme survivor” species to thrive in the nitrogen-poor soils of Chile’s hyper-arid Atacama Desert. Our collections focus on 32 Atacama extremophile species, including seven grass species with potential biofuel applications. To identify genes-of-importance to survival we compared genomic and transcriptomic profiles of extremophile species that thrive in the Atacama to those of closely related “sister” species from nitrogen-rich arid and mesic regions of California. Deep RNA sequencing and de novo transcriptome assembly across these triplet species sets supported a phylogenomic framework for identifying positively selected genes associated with adaptive divergence. Our integrative analysis combined ecological and environmental data, metagenomics, evolutionary and systems biology, and metabolomics. This enabled us to create an unprecedented framework for systematically understanding how non-model plants have adapted to survive in extreme conditions. Our resulting database of positively selected ortholog groups in the extremophile plants offers promising targets for engineering crop and biofuel species with enhanced resilience to drought and extreme weather. Additionally, our newest dataset explores and exploits a complementary metabolomic approach. This new aspect provides innovative strategies to manipulate plant cell metabolism, further supporting efforts to improve agricultural productivity in the face of extreme climates. Importantly, our combined evolutionary- and metabolomic-based strategies focused on convergent patterns of adaptation, providing a genetic and metabolomic toolkit for improving crop and biofuel resilience across diverse plant species. Finally, our novel exploration of ecological and evolutionary dynamics delivered to the community a phylogenomic computational pipeline called “PhyloGeneious.” Our continued adaptations of this pipeline are publicly available to expedite evolutionary genomic research for future scientific discoveries. In total, our DOE-BER has provided genomic, metabolomic, and computational strategies to understand how extremophile plants provide evolutionary and physiological targets for improving agricultural and biofuel production.

59 BASIC BIOLOGICAL SCIENCES↗

Artificial intelligence–powered biofoundries for protein engineering and metabolic engineering

Synthetic biology is rapidly evolving through the integration of artificial intelligence (AI) and automated biofoundries. This convergence accelerates the design–build–test–learn cycle, shifting protein engineering and metabolic engineering from labor-intensive manual experimentation to autonomous experimentation. This review summarizes recent advances in workflow development, AI models, and their integration with biofoundries for automated or autonomous protein engineering and metabolic engineering. Particularly, we highlight the potential of AI-powered biofoundries for accelerated scientific discovery and innovation in synthetic biology.

Chen, Junyu [Univ. of Illinois at Urbana-Champaign↗

Phenylpropanoid methyl esterase unlocks catabolism of aromatic biological nitrification inhibitors

Microbial nitrification of fertilizers represents is a significant global source of greenhouse gas emissions. This process increases emissions, fosters toxic algal blooms, and raises crop production costs. Some plants naturally release biological nitrification inhibitors to suppress ammonium-oxidizing microbes and reduce nitrification. Engineering nitrification inhibitor production into food and bioenergy crops via synthetic biology offers a promising mitigation strategy, but its success depends on addressing gaps in our understanding of inhibitor degradation in soil. This study begins to fill this gap by identifying a previously unknown microbial pathway for degrading phenylpropanoid methyl esters, a key class of aromatic nitrification inhibitors. Using transcriptomics and high-throughput functional genomics, we discovered genes essential for phenylpropanoid methyl ester degradation. Genetic and biochemical analyses revealed two novel enzymes, including a newly identified phenylpropanoid methyl esterase, that direct phenylpropanoid methyl esters into known metabolic pathways. Importantly, transferring these genes into bacteria capable of metabolizing other phenylpropanoids enabled them to use the methyl esters as a carbon source. This work provides critical insights into microbial nitrification inhibitor degradation, a poorly understood element of the nitrification cycle.

Genetic Engineering↗

The Artificial Intelligence Ontology: LLM-Assisted Construction of AI Concept Hierarchies

The Artificial Intelligence Ontology (AIO) is a systematization of artificial intelligence (AI) concepts, methodologies, and their interrelations. Developed via manual curation, with the additional assistance of large language models (LLMs), AIO aims to address the rapidly evolving landscape of AI by providing a comprehensive framework that encompasses both technical and ethical aspects of AI technologies. The primary audience for AIO includes AI researchers, developers, and educators seeking standardized terminology and concepts within the AI domain. We use the term “branches” for classes, and their subclasses, in our ontology that are subclasses of owl:Thing. AIO contains eight branches: Bias, Layer, Machine Learning Task, Mathematical Function, Model, Network, Preprocessing, and Training Strategy, each designed to support the modular composition of AI methods and facilitate a deeper understanding of deep learning architectures and ethical considerations in AI. AIO uses the Ontology Development Kit (ODK) for its creation and maintenance, with its content being more easily updated through AI-driven curation support. This approach not only ensures the ontology's relevance amidst the fast-paced advancements in AI but also significantly enhances its utility for researchers, developers, and educators by simplifying the integration of new AI concepts and methodologies. The ontology's utility is demonstrated through the annotation of AI methods data in a catalog of AI research publications and the integration into the BioPortal ontology resource, highlighting its potential for cross-disciplinary research. The AIO ontology is open source and is available on GitHub ( https://w3id.org/aio/ ) and BioPortal ( https://bioportal.bioontology.org/ontologies/AIO ).

Joachimiak, Marcin P. [Biosystems Data Science Dep↗

Pooled PPIseq: Screening the SARS-CoV-2 and human interface with a scalable multiplexed protein-protein interaction assay platform

Protein-Protein Interactions (PPIs) are a key interface between virus and host, and these interactions are important to both viral reprogramming of the host and to host restriction of viral infection. In particular, viral-host PPI networks can be used to further our understanding of the molecular mechanisms of tissue specificity, host range, and virulence. At higher scales, viral-host PPI screening could also be used to screen for small-molecule antivirals that interfere with essential viral-host interactions, or to explore how the PPI networks between interacting viral and host genomes co-evolve. Current high-throughput PPI assays have screened entire viral-host PPI networks. However, these studies are time consuming, often require specialized equipment, and are difficult to further scale. Here, we develop methods that make larger-scale viral-host PPI screening more accessible. This approach combines the mDHFR split-tag reporter with the iSeq2 interaction-barcoding system to permit massively-multiplexed PPI quantification by simple pooled engineering of barcoded constructs, integration of these constructs into budding yeast, and fitness measurements by pooled cell competitions and barcode-sequencing. We applied this method to screen for PPIs between SARS-CoV-2 proteins and human proteins, screening in triplicate >180,000 ORF-ORF combinations represented by >1,000,000 barcoded lineages. Our results complement previous screens by identifying 74 putative PPIs, including interactions between ORF7A with the taste receptors TAS2R41 and TAS2R7, and between NSP4 with the transmembrane KDELR2 and KDELR3. We show that this PPI screening method is highly scalable, enabling larger studies aimed at generating a broad understanding of how viral effector proteins converge on cellular targets to effect replication.

60 APPLIED LIFE SCIENCES↗

NuclPred v1

This tool takes a genome assembly as input and predicts per-site nucleosome occupancy as output. Trained on physical maps of nucleosome binding preferences across the fungal kingdom, NuclPred can be applied broadly across fungi (and other eukaryotes). This breadth, combined with its accuracy, means it could have both basic and applied biological implications, for example in understanding eukaryotic gene regulation and genetic engineering. Almost universally across eukaryotes, nucleosomes - each wrapping ~150 base pairs of DNA - serve to package DNA inside the nucleus, with major consequences on DNA access, gene activity and DNA integration. NuclPred was generated using a supervised deep learning approach combining convolutional and recurrent neural networks to take DNA features (nucleotides, GC content and structural information) as input, then use that information to predict the physical attractiveness DNA sequences might have for forming nucleosomes. With this information at hand, researchers can design more efficient CRISPR constructs, explore the interplay between DNA signatures and other regulators impact nucleosome locations, predict expression patterns, etc. This tool will be published as part of a manuscript currently under revision at iScience (draft attached).

Mondo, Stephen↗