Search NASASearch

SEARCH · Search NASA

Results for “DNA construct”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

New principles of self‐organization created through the interplay of DNA condensates, microtubules, and motors

Bioinspired design—which holds great promise for a new generation of materials that are robust to defects, scalable under green manufacture, environmentally responsive, and programmably reconfigurable—requires mastery over molecular self-organization. Yet, from its specific mechanisms to most general architectures, the principles governing self-organization remain poorly understood and not even fully enumerated. For living systems, one obvious architectural principle is the modular reuse of a few simple molecular components in myriad combinations to achieve more complex phenomena. For example, the mechanical tasks of a cell are driven by the nonequilibrium dynamics of cytoskeletal filaments and molecular motors—the same filaments and motors, reprogrammed by a variety of modulators, perform tasks ranging from cell movement to division. Similarly, many compartmentalization tasks are performed by liquid-like condensates of simple components, which act as membraneless organelles to localize particular molecules in space and time (e.g. for gene regulation or RNA processing). In a few cases, condensates combine and interact with the cytoskeleton to create still more complex phenomena, e.g. the nucleation of microtubule asters from the centrosome (a protein condensate) to form the mitotic spindle during cell division. Very little is known about the fundamental mechanisms of such filament-plus-condensate phenomena. Despite few examples, the landscape of behaviors that can be achieved through the combination of condensates, filaments, and motors appears vast. However, exploration has been hindered by a lack of systems that have sufficiently programmable and dynamically tunable interactions between component condensates, filaments, and motors. We proposed to combine programmable DNA condensates, filamentous microtubules, and light-controlled motors into self-organizing systems whose principles go beyond those that have been observed in nature. In one limit, our systems will use microtubules and motors to create the molecular analog of a network of roads, which will organize droplets of DNA condensates capable of carrying molecular cargo. DNA condensates coupled to motors will flow from one microtubule aster hub to another, with their direction and timing controlled by DNA circuits. In another limit, microtubules will swim through bulk DNA condensates and exhibit strong interactions with boundaries between different types of condensates. Microtubule swimmers will reflect, get trapped, or refract at boundaries, under a mechanical analog of the classical optical index of refraction. DNA condensates having different mechanical indexes of refraction will be used to construct the analog of optical lenses, so that microtubule swimmers can be manipulated like light—collimated, diffracted, focused, and sorted based on properties analogous to wavelength. These two limits define two new architectures, within which multiple new mechanistic principles for self-organization will be discovered and explored. To explore these architectures, the motor-based coupling between DNA condensates and filaments will be controlled in time and space through the use of opto-proteins that create reversible links between DNA condensates and motors upon illumination. For each principle of interest, patterns of light will create virtual experiments by defining patterns of activity where DNA condensates walk along filaments, or filaments swim through condensates, and patterns of inactivity which will serve either as controls, or as boundary conditions vital to create the desired phenomena. This research serves the goals of Basic Energy Sciences Biomolecular Material Program by elucidating the principles by which the emergent, nonequilibrium behavior of collections of DNA condensates, motors, and microtubules can be programmed by environmental light patterns to create complex motion and materials transport. Because DNA condensates can be readily coupled to virtually any high performance nanomaterial, from carbon nanotubes, to metal nanoparticles, to light harvesting systems, this work provides a path to the construction, self-maintenance and reconfiguration of materials relevant to the Department of Energy.

60 APPLIED LIFE SCIENCES

Harnessing Heterologous Bacterial Two-Component Systems as Biosensors to Address Challenges in Fermentation Scale-Up

Scaling up bacterial fermentation from bench to industrial scale often results in unpredictable performance losses, possibly in part due to changes in microenvironmental conditions such as pH. To investigate this, we developed a suite of pH-sensitive biosensors from bacterial two-component systems (TCSs) that provide a dynamic, fluorescent readout in response to extracellular pH changes. TCSs consist of a periplasmic sensor histidine kinase (HK) that, in response to an extracellular stimulus, autophosphorylates intracellularly and subsequently transfers the phosphate to a cognate response regulator (RR) that modulates transcription of target genes. We utilized three pH-responsive TCSs (referred to here as CVJ1, CVJ30, and CVJ79) and linked their output to GFP. This was achieved by placing the RR promoter upstream of GFP or by constructing a chimeric RR composed of the native receiver domain and the DNA-binding domain of another well-characterized RR with a defined promoter. All components - HK, RR (native or chimeric), and GFP under its corresponding promoter - were cloned into a broad-host-range plasmid. Sensors were validated in Escherichia coli and Pseudomonas putida, including the muconic acid-producing strain P. putida TL207. All three biosensors successfully reported pH, with fluorescence (normalized to optical density) correlating strongly with media pH. Among the native sensors, CVJ79 showed the most robust performance while CVJ1 also performed best in its native form; CVJ30 exhibited improved functionality as a chimera, suggesting that modular RR design can enhance compatibility in some heterologous hosts. Further, CVJ79 was activated by alkaline conditions, while CVJ30 responded to acidic environments. Notably, CVJ1 was induced by high pH in wild-type E. coli and P. putida, but low pH in TL207. The observed differences in sensor activation between strains - particularly the divergent response of CVJ1 - suggest that host-specific regulatory pathways may influence how cells perceive and adapt to pH stress. Moving forward, these biosensors can be used to guide the rational design of more robust strains, optimize process conditions in real time, and inform strategies to minimize physiological heterogeneity during scale-up. Integrating these tools into high-throughput screening and bioreactors will be a key step toward improving predictability and performance in industrial bioprocesses.

09 BIOMASS FUELS

Thermophilic site-specific recombination system for rapid insertion of heterologous DNA into the Clostridium thermocellum chromosome

Clostridium thermocellum is an anaerobic thermophile capable of producing ethanol and other commodity chemicals from lignocellulosic biomass. The insertion of heterologous DNA into the C. thermocellum chromosome is currently achieved via a time-consuming homologous recombination process, where a single stable insertion can take 2–4 weeks or more to construct. In this work, we developed a thermostable version of the Serine recombinase Assisted Genome Engineering (tSAGE) approach for gene insertion in C. thermocellum utilizing a site-specific recombinase from Geobacillus sp. Y412MC61, enabling quick and easy insertion of DNA into the chromosome for accelerated genetic tool screening and heterologous gene expression. Using tSAGE, chromosomal insertion of plasmid DNA occurred at a maximum transformation efficiency of 5 × 10 3 CFU/µg, which is comparable to the transformation efficiency of a replicating control plasmid in C. thermocellum. Using tSAGE, we chromosomally integrated and characterized 17 reporter genes, 15 homologous and 31 heterologous constitutive promoters of varying strengths, 4 inducible promoters, and 5 riboswitches in C. thermocellum. We also determined that a 6–7 nucleotide gap between the ribosome binding site (RBS) and the start codon is optimal for high expression by employing a library of superfolder green fluorescent protein expression constructs driven by our strongest tested promoter (P clo1313_1194 ) with different distances between the RBS and start codon. The tools developed here will aid in accelerating C. thermocellum strain engineering for producing sustainable fuels and chemicals directly from plant biomass.

Biofuels

Developmental assembly of multi-component polymer systems through interconnected synthetic gene networks in vitro

Abstract Living cells regulate the dynamics of developmental events through interconnected signaling systems that activate and deactivate inert precursors. This suggests that similarly, synthetic biomaterials could be designed to develop over time by using chemical reaction networks to regulate the availability of assembling components. Here we demonstrate how the sequential activation or deactivation of distinct DNA building blocks can be modularly coordinated to form distinct populations of self-assembling polymers using a transcriptional signaling cascade of synthetic genes. Our building blocks are DNA tiles that polymerize into nanotubes, and whose assembly can be controlled by RNA molecules produced by synthetic genes that target the tile interaction domains. To achieve different RNA production rates, we use a strategy based on promoter “nicking” and strand displacement. By changing the way the genes are cascaded and the RNA levels, we demonstrate that we can obtain spatially and temporally different outcomes in nanotube assembly, including random DNA polymers, block polymers, and as well as distinct autonomous formation and dissolution of distinct polymer populations. Our work demonstrates a way to construct autonomous supramolecular materials whose properties depend on the timing of molecular instructions for self-assembly, and can be immediately extended to a variety of other nucleic acid circuits and assemblies.

Science & Technology - Other Topics

A multi-omic characterization of the physiological responses to salt stress in Scenedesmus obliquus UTEX393

Scenedesmus obliquus UTEX393 is a promising microalgal candidate for sustainable biomanufacturing but its limited halotolerance hinders large-scale cultivation in saline environments. To investigate the molecular basis of salt stress responses, we conducted a comprehensive multi-omic analysis integrating genomics, transcriptomics, proteomics, lipidomics, metabolomics, and DNA affinity purification sequencing (DAP-seq). An improved nuclear genome assembly and annotation yielded 19,017 gene models and a 97% BUSCO completeness score, enabling construction of a genome-scale metabolic model. Comparing 15 ppt salinity stress to 5 ppt control, growth and productivity were significantly reduced, accompanied by widespread transcriptomic and proteomic changes. Transcriptomic analysis revealed downregulation of photosynthetic machinery and energy conservation genes, and upregulation of stress-responsive elements such as expansins, flavodoxins, and osmoprotectants. Lipidomic profiling showed accumulation of triacylglycerols (TAGs) and degradation of galactosyl lipids, consistent with a shift toward lipid biosynthesis to mitigate redox imbalance. Depletion of key polar metabolites and branched-chain amino acids suggested a rerouting of central carbon metabolism under stress. DAP-seq identified key transcription factors, including LHY1 and SPL12, that target central metabolic enzymes involved in redox balancing, such as glyceraldehyde-3-phosphate dehydrogenase (GAPDH) and malate dehydrogenase (MDH). These findings establish a regulatory-metabolic framework linking redox stress to lipid accumulation and reveal potential engineering targets to enhance salt tolerance. Overall, the multi-omic analysis supports the “overflow” hypothesis, where impaired photosynthesis results in excess reducing equivalents being diverted into TAG synthesis and highlights transcriptional regulators as candidates for improving algal robustness in brackish environments.

09 BIOMASS FUELS

Pooled PPIseq: Screening the SARS-CoV-2 and human interface with a scalable multiplexed protein-protein interaction assay platform

Protein-Protein Interactions (PPIs) are a key interface between virus and host, and these interactions are important to both viral reprogramming of the host and to host restriction of viral infection. In particular, viral-host PPI networks can be used to further our understanding of the molecular mechanisms of tissue specificity, host range, and virulence. At higher scales, viral-host PPI screening could also be used to screen for small-molecule antivirals that interfere with essential viral-host interactions, or to explore how the PPI networks between interacting viral and host genomes co-evolve. Current high-throughput PPI assays have screened entire viral-host PPI networks. However, these studies are time consuming, often require specialized equipment, and are difficult to further scale. Here, we develop methods that make larger-scale viral-host PPI screening more accessible. This approach combines the mDHFR split-tag reporter with the iSeq2 interaction-barcoding system to permit massively-multiplexed PPI quantification by simple pooled engineering of barcoded constructs, integration of these constructs into budding yeast, and fitness measurements by pooled cell competitions and barcode-sequencing. We applied this method to screen for PPIs between SARS-CoV-2 proteins and human proteins, screening in triplicate >180,000 ORF-ORF combinations represented by >1,000,000 barcoded lineages. Our results complement previous screens by identifying 74 putative PPIs, including interactions between ORF7A with the taste receptors TAS2R41 and TAS2R7, and between NSP4 with the transmembrane KDELR2 and KDELR3. We show that this PPI screening method is highly scalable, enabling larger studies aimed at generating a broad understanding of how viral effector proteins converge on cellular targets to effect replication.

60 APPLIED LIFE SCIENCES

Through the lens of bioenergy crops: advances, bottlenecks, and promises of plant engineering

Advances in engineering of bioenergy crops were driven over the past years by adapting technological breakthroughs and accelerating conventional applications but also exposed intriguing challenges. New tools revealed rich interconnectivity in the exponentially growing and dynamic 'big' omics data' of metabolomes, transcriptomes, and genomes at previously inaccessible magnitude (global, cross-species, meta-) and resolution (single cell). Insights enabled fresh hypotheses and stimulated disciplines such as functional genomics with discovery of broad regulatory networks and their determinants, that is, DNA parts, including promoters, regulatory elements, and transcription factors. Their rational design, assembly into increasingly complex blueprints, and installation into diverse chassis is an existing frontier that may benefit from emerging technologies to address bottlenecks. Interweaving nature-inspired to fully synthetic parts has already allowed building of fine-tuned regulatory circuits, or new-to-nature metabolic routes insulated from the biological context of the chassis species. Similarly, developments and the evolving need for unifying principles in plant transformation and species-agnostic technologies highlight future opportunities for engineering the next generation of bioenergy plants.

60 APPLIED LIFE SCIENCES

A call to standardize metrics for monitoring baleen whales near marine construction activities

Effective monitoring is necessary to protect marine mammal species during the construction of offshore infrastructure. The tools for detecting or monitoring marine mammals span traditional (e.g., visual observers, optical cameras), to newer (e.g., passive acoustic monitoring, infrared cameras, tags), and emerging (e.g., satellite imagery, environmental DNA, dimethyl sulfide concentration) technologies. Some are better suited for use during offshore development; however, peer-reviewed literature does not typically evaluate and report on the performance of these various technologies. We define a minimum set of metrics related to efficacy (i.e., confusion matrix, precision and recall, probability of missed mitigation), detection range (i.e., maximum and reliable detection range, spatial resolution), and data delivery (i.e., detection latency, system reliability, temporal resolution) that we recommend are needed to assess the utility of monitoring technologies for this purpose. Following a literature review of relevant studies, we highlight which publications reported these metrics and used multiple technologies to compare relative performance. We also emphasize the benefits of multi-modal approaches and recommend performance assessments through modeling or large-scale collaborative field testing. These metrics will standardize data collection, reporting, and analysis; promote consistent and comparable results; and foster collaboration among developers, regulatory agencies, and scientists. This may lead to the co-development of technology that achieves multiple goals, has greater application, and can answer research questions while collecting data to fulfill permitting requirements. These metrics may also inform decisions on what systems regulatory agencies might consider using and reduce monitoring costs, which is critical to support the marine sector's rapid growth alongside marine mammal conservation.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI

Complete replacement of Arabidopsis oil-producing enzymes with heterologous diacylglycerol acyltransferases

Acyl-CoA:diacylglycerol acyltransferase 1 (DGAT1) and phospholipid:diacylglycerol acyltransferase 1 (PDAT1) share responsibility for triacylglycerol (TAG) biosynthesis, and their selectivities control TAG fatty acid (FA) compositions. For rational metabolic engineering of seed oils, replacing endogenous TAG biosynthesis with exogenous enzymes containing different substrate FA selectivities is desirable; however, the dgat1-1/pdat1-2 double mutant is pollen lethal. Here, we evaluated the ability of 3 DGAT1s, from phylogenetically diverse plants with distinct TAG assembly processes, to completely replace endogenous TAG biosynthesis in Arabidopsis ( Arabidopsis thaliana ). We transformed dgat1-1 mutant plants with expression constructs for DGAT1 s from Camelina sativa , Physaria fendleri , and castor ( Ricinus communis ). Transgene expression was properly “contextualized” by using a previously determined minimum necessary expression unit containing the promoter/5′ UTR and first intron of native AtDGAT1 ; both of these DNA elements are essential for pollen expression. Next, we crossed homozygous lines with a DGAT1/DGAT1/PDAT1/pdat1-2 parent. C. sativa and P. fendleri DGAT1s restored the FA compositions and transcriptional differences of dgat1-1 to near wild-type and rescued the dgat1-1/pdat1-2 pollen lethality. R. communis DGAT1 was active in dgat1-1 seeds but produced unique oil profiles and alterations in the expression of lipid metabolic genes; it also failed to rescue dgat1-1/pdat1-2 lethality. This study confirms that the promoter and first intron of AtDGAT1 can modulate the expression of foreign DGAT1 genes to fit the correct spatiotemporal profile necessary for completely replacing endogenous TAG biosynthesis. Furthermore, it demonstrates an additional layer of unexpected enzyme incompatibility between oilseed lineages, which may complicate bioengineering approaches that seek to replace essential genes with orthologs.

McGuire, Sean T. [Washington State Univ., Pullman,

Data for Protoplast Fusion as a Strategy to Increase Ploidy in Rhodotorula toruloides for Strain Development

Rhodotorula toruloides is a red oleaginous yeast with growing commercial interest because of its hardiness and exceptional lipid production capacity. Because it is a basidiomycete yeast with a complex life cycle, many of the classical breeding methods used with ascomycetes are unavailable for strain improvement. However, we have been able to construct polyploid yeast by fusing protoplasts of parents with the same mating type. Fusing of Y-6985 (A2) and Y-48190 (A2), which had been transformed with complementary antibiotic markers, led to the recovery of two diploids and one triploid. The stability of the fusion yeasts was tested by plating them on non-selective medium after several growth cycles under antibiotics and then testing five colonies per strain for nuclear DNA contents using flow cytometry and standard cell cycle analysis: the triploid and one diploid were stable. Fusants inherited their mitochondria from a single parent, which was demonstrated using restriction fragment length polymorphism (RFLP) of mitochondrial DNA. The phenotypic properties of the parents and fusants were compared in glucose fed-batch bioreactor studies and cellulosic sugar batch cultures. The final lipid titers for the fed-batch cultures were 24.9–39.7 g/L with Y-6985 and the diploid and triploid performing the best and worst, respectively. The fusants demonstrated intermediate hardiness for growth on hydrolysate prepared with dilute-acid pretreated switchgrass and were outperformed by Y-48190. Unlike one of the haploid parents, the fusants grew in 70% v/v concentrated hydrolysate. However, they did not grow as fast as the other haploid. In this study, a modernized protoplast fusion method is resurrected a useful tool for strain development in this yeast, which is complementary with other available methods.

FOS: Biological sciences

Models and Algorithms for Equilibrium Analysis of Mixed-Material Nucleic Acid Systems

Dynamic programming algorithms within the NUPACK software suite enable analysis of equilibrium base-pairing properties for complex and test tube ensembles containing arbitrary numbers of interacting nucleic acid strands. Currently, calculations are limited to single-material systems that are either all-RNA or all-DNA. Here, to enable analysis of mixed-material systems that are critical for modern applications in vitro, in situ, and in vivo, we develop physical models and dynamic programming algorithms that allow the material of the system to be specified at nucleotide resolution. Free energy parameter sets are constructed for both RNA/DNA and RNA/2'OMe-RNA mixed-material systems by combining available empirical mixed-material parameters with single-material parameter sets to enable treatment of the full complex and test tube ensembles. New dynamic programming recursions account for the material of each nucleotide throughout the recursive process. For a complex with N nucleotides, the mixed-material dynamic programming algorithms maintain the O(N 3 ) time complexity of the single-material algorithms, enabling efficient calculation of diverse physical quantities over complex and test tube ensembles (e.g., complex partition function, equilibrium complex concentrations, equilibrium base-pairing probabilities, minimum free energy secondary structure(s), and Boltzmann-sampled secondary structures) at a cost increase of roughly 2.0-3.5×. The results of existing single-material algorithms are exactly reproduced when applying the new mixed-material algorithms to single-material systems. Accuracy is significantly enhanced using mixed-material models and algorithms to predict RNA/DNA and RNA/2'OMe-RNA duplex melting temperatures from the experimental literature as well as RNA/DNA melt profiles from new experiments. In conclusion, mixed-material analyses can be performed online using the NUPACK web app (www.nupack.org) or locally using the NUPACK Python module.

2′OMe-RNA

Nanopore Activity Assays for Detection of Biomarker Protease Activity: Design and Testing of Substrates for Both Nanopore Sequencing and PCR-Based Detection Methods

The work performed in this project has demonstrated the ability to construct proteolytic enzyme substrates that are PCR and sequencing-readable reporter molecules. Specifically, the goal was to detect those reporter molecules via PCR and Oxford Nanopore Technologies MinION sequencing methods following exposure to the biomarker protease thrombin. The assay development focused on binding the constructed peptide-oligonucleotide chimera to immobilized streptavidin. The action of thrombin on the peptide portion of the molecule released the oligonucleotide for detection. Detection of protease activity was demonstrated in a concentration-dependent manner using MALDI-MS, RT-PCR and DNA sequencing. Additional steps to remove background release of reporter molecules during the assay was used to improve the difference in detected oligonucleotide reporter following protease activity. Additional steps in assay development will be to (1) test the assay in an appropriate matrix, (2) investigate detection using additional DNA sequencing platforms and (3) demonstrate multiplexed detection of multiple protease markers in a single reaction.

59 BASIC BIOLOGICAL SCIENCES

Nucleic Acid-Based Detection Protease Activity

Proteases include clinically relevant markers for clotting disorders, certain cancers as well as toxins. Assays for protease activity often use designed peptides mimicking natural substrates and detection with colorometric and fluorescence-based detection that is difficult to multiplex without expensive and resource demanding instruments. This work demonstrates detection of proteolytic activity using PCR and sequencing-readable reporter molecules. The assay development focused on binding the constructed peptide-oligonucleotide chimera to immobilized streptavidin. Thrombin, an essential component of the clotting cascade, was used as a model system for testing peptide substrate recognition and release of a designed oligonucleotide for detection. Detection of protease activity was demonstrated in a concentration-dependent manner using MALDI-MS, RT-PCR and DNA sequencing.

Wunschel, David S [Pacific Northwest National Labo

Yeast Transformation on Hamilton Vantage (YT Vantage) v1

Our software program is designed for the Hamilton Vantage liquid handling robot, automating the Build step in the Design-Build-Test-Learn (DBTL) cycle for Saccharomyces cerevisiae. This program minimizes human intervention, enabling rapid identification of pathway bottlenecks and genes that enhance verazine production. The program takes competent yeast and plasmid DNA as input and generates an output library of engineered strains compatible with automated colony picking, high-throughput culturing, and chemical extraction for downstream LC-MS analysis. A user-friendly interface, developed using the Hamilton Method Editor software, allows for on-demand parameter customization. By automating this process, our program streamlines the construction of Saccharomyces cerevisiae, reducing manual labor and increasing efficiency. While the manual process is well-documented, integration with robotic automation is less common, making our program a valuable tool for researchers. With this software, we achieved 2-5 fold increases in verazine production, demonstrating its potential to accelerate research in this field.

Louie, Randy [Lawrence Berkeley National Laborator

Protoplast fusion as a strategy to increase ploidy in Rhodotorula toruloides for strain development

Rhodotorula toruloides is a red oleaginous yeast with growing commercial interest because of its hardiness and exceptional lipid production capacity. Because it is a basidiomycete yeast with a complex life cycle, many of the classical breeding methods used with ascomycetes are unavailable for strain improvement. However, we have been able to construct polyploid yeast by fusing protoplasts of parents with the same mating type. Fusing of Y-6985 (A2) and Y-48190 (A2), which had been transformed with complementary antibiotic markers, led to the recovery of two diploids and one triploid. The stability of the fusion yeasts was tested by plating them on non-selective medium after several growth cycles under antibiotics and then testing five colonies per strain for nuclear DNA contents using flow cytometry and standard cell cycle analysis: the triploid and one diploid were stable. Fusants inherited their mitochondria from a single parent, which was demonstrated using restriction fragment length polymorphism (RFLP) of mitochondrial DNA. The phenotypic properties of the parents and fusants were compared in glucose fed-batch bioreactor studies and cellulosic sugar batch cultures. The final lipid titers for the fed-batch cultures were 24.9–39.7 g/L with Y-6985 and the diploid and triploid performing the best and worst, respectively. The fusants demonstrated intermediate hardiness for growth on hydrolysate prepared with dilute-acid pretreated switchgrass and were outperformed by Y-48190. Unlike one of the haploid parents, the fusants grew in 70% v/v concentrated hydrolysate. Furthermore, they did not grow as fast as the other haploid. In this study, a modernized protoplast fusion method is resurrected a useful tool for strain development in this yeast, which is complementary with other available methods.

Lignocellulose

Structural and compositional complexities of hierarchical self-assembly: A hypergraph approach

Programmable self-assembly enables the construction of complex molecular, supramolecular, and crystalline architectures from well-designed building blocks. In this work, we introduce a hypergraph-based formalism, Blocks & Bonds (B&B), which generalizes classical chemical graph theory by incorporating directed and multicolored interactions, internal symmetries, and hierarchical organization. Within this framework, we develop the Structure Code (SC), a compact and versatile language for describing self-assembled architectures. We define a Kolmogorov-style structural complexity as the total information content of SC, obtained through its tokenization and Shannon information assignment. Complementing this encoding-based measure, we introduce a much simpler quantity, the compositional complexity, which depends only on the number and cumulative usage of block and bond types in the construction set. A central result of this work is a strong empirical correlation between the token-based structural complexity and the compositional complexity across all examined systems. Owing to this agreement, the compositional complexity emerges as the most practical and broadly applicable measure: it is easy to compute, requires no explicit encoding, and yet closely tracks the actual information content of structurally diverse architectures. Applications to molecular systems (ethylene glycol and glucose), DNA-origami lattices, and crystalline assemblies show that B&B hypergraphs provide a unified, scalable, and information-efficient representation of structural organization, naturally capturing symmetry, modularity, and stereochemistry. This framework establishes a quantitative foundation for complexity-aware classification and inverse design of programmable matter.

36 MATERIALS SCIENCE

Expanding the genetic toolkit: adenine and cytosine base editors for gene disruption in Aspergillus niger

Despite revolutionizing fungal genetic engineering, conventional CRISPR/Cas9-mediated knockouts rely on DNA double-strand breaks (DSBs), which can cause unwanted insertions and deletions, chromosomal abnormalities, and cytotoxicity. Base editors such as adenine base editors (ABEs), which convert A‧T to G‧C, and cytosine base editors (CBEs), which convert C‧G to T‧A, offer a safer alternative by enabling predictable, target-specific single-nucleotide changes without introducing DSBs. To overcome the limitations of traditional genome editing in filamentous fungi, we developed efficient base-editing systems in Aspergillus niger . For the first time, we constructed an ABE in A. niger , achieving up to 80% editing efficiency and inducing predictable A-to-G mutations at the intended intron sites, disrupting gene function through mRNA mis-splicing. We also developed a highly efficient CBE system, capable of introducing premature stop codons with 50–100% efficiency. To broaden the editing scope, we implemented a Cas9-NG variant recognizing a relaxed PAM sequence requiring only a single guanine (G), enabling editing at start codons and splice sites. Leveraging this expanded scope, we established gene disruption approaches by targeting start codons via ABE-mediated A-to-G conversions (ATG-to-GTG and ATG-to-ACG) and CBE-mediated C-to-T conversion (ATG-to-ATA). Additionally, our base-editing systems enable multiplex gRNA delivery and marker-free editing of multiple genes. Collectively, the scope-expanding strategies increase the number of genes targetable for disruption by base-editing in A. niger by 26.3% and enable near-complete coverage of 96% of the coding genes. Overall, this work demonstrates the potential of ABE and CBE systems as versatile, efficient, and safer alternatives to DSBs-based gene disruption in filamentous fungi.

Aspergillus

Finding the missing pieces: filling gaps that impede the translation of omics data into models

High-throughput omics technologies such as DNA sequencing have made the sequencing and computational assembly of microbial genomes recovered from the environment relatively routine. Computational inference of the protein products encoded by these genomes, and the associated biochemical functions, should enable the accurate prediction and modeling of microbial metabolism, organismal interactions, and ecosystem processes. However, a lack of scalable, probabilistic protein annotation tools limits the full potential of modeling for understanding the metabolism and biogeochemical cycles of microbial communities. Our approach to improve inference of protein annotations and metabolic models relied on learning from and emulating expert manual curation, leveraging software engineering and data science best practices to scale up the throughput and accuracy of annotations and metabolic model construction, building software to objectively evaluate different annotation strategies, and more closely linking the protein annotation and metabolic model inference process. Outcomes of this research include several improved or new computational tools, including DRAM (Distilled and Refined Annotation of Metabolism) for annotating microbial genomes with protein function and metabolic traits, CAMPER (Curated Annotations for Microbial Polyphenol Enzymes and Reactions) for annotating key polyphenol metabolisms, EC-Bench for comprehensive and unbiased benchmarking of annotation tools, and several apps available via the DOE Systems Biology Knowledgebase (KBase) for building genome-scale metabolic models. We demonstrate that these tools allow us to scalably annotate and understand thousands of genomes for microbial communities from a variety of systems and test cases, including rivers, thawing permafrost, and gut microbiomes. All of these computational tools are available as open-source software, with most broadly and easily accessible to the scientific community via KBase apps.

59 BASIC BIOLOGICAL SCIENCES