Search NASA⌕ Search

SEARCH · Search NASA

Results for “Conserved Sequence”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Next-Generation Sequencing Data from a CUT&RUN Study of R. toruloides IFO0880 Cse4 and Orc1 Binding Sites

Rhodotorula toruloides has been increasingly explored as a host for bioproduction of lipids, fatty acid derivatives and terpenoids. Various genetic tools have been developed, but neither a centromere nor an autonomously replicating sequence (ARS), both necessary elements for stable episomal plasmid maintenance, has yet been reported. In this study, cleavage under targets and release using nuclease (CUT&RUN), a method used for genome-wide mapping of DNA–protein interactions, was used to identify R. toruloides IFO0880 genomic regions associated with the centromeric histone H3 protein Cse4, a marker of centromeric DNA. Fifteen putative centromeres ranging from 8 to 19 kb in length were identified and analyzed, and four were tested for, but did not show, ARS activity. These centromeric sequences contained below average GC content, corresponded to transcriptional cold spots, were primarily nonrepetitive and shared some vestigial transposon-related sequences but otherwise did not show significant sequence conservation. Future efforts to identify an ARS in this yeast can utilize these centromeric DNA sequences to improve the stability of episomal plasmids derived from putative ARS elements.

Genome Engineering↗

Characterization of lignin-degrading enzyme PmdC, which catalyzes a key step in the synthesis of polymer precursor 2-pyrone-4,6-dicarboxylic acid

Pyrone-2,4-dicarboxylic acid (PDC) is a valuable polymer precursor that can be derived from the microbial degradation of lignin. The key enzyme in the microbial production of PDC is 4-carboxy-2-hydroxymuconate-6-semialdehyde (CHMS) dehydrogenase, which acts on the substrate CHMS. We present the crystal structure of CHMS dehydrogenase (PmdC from Comamonas testosteroni) bound to the cofactor NADP, shedding light on its three-dimensional architecture, and revealing residues responsible for binding NADP. Using a combination of structural homology, molecular docking, and quantum chemistry calculations, we have predicted the binding site of CHMS. Key histidine residues in a conserved sequence are identified as crucial for binding the hydroxyl group of CHMS and facilitating dehydrogenation with NADP. Mutating these histidine residues results in a loss of enzyme activity, leading to a proposed model for the enzyme's mechanism. These findings are expected to help guide efforts in protein and metabolic engineering to enhance PDC yields in biological routes to polymer feedstock synthesis.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Utilizing Machine Learning to Improve Neutralization Potency of an HIV-1 Antibody Targeting the gp41 N-Heptad Repeat

The N-heptad repeat (NHR) of the HIV-1 gp41 prehairpin intermediate (PHI) is an attractive potential vaccine target with high sequence conservation across diverse strains. However, despite the potency of NHR-targeting peptides and clinical efficacy of the NHR-targeting entry inhibitor enfuvirtide, no potently neutralizing NHR-directed monoclonal antibodies (mAbs) nor antisera have been identified or elicited to date. The lack of potent NHR-binding mAbs both dampens enthusiasm for vaccine development efforts at this target and presents a barrier to performing passive immunization experiments with NHR-targeting antibodies. To address this challenge, we previously developed an improved variant of the NHR-directed mAb D5, called D5_AR, which is capable of neutralizing diverse tier-2 viruses. Building on that work, here we present the 2.7Å-crystal structure of D5_AR bound to NHR mimetic peptide IQN17. We then utilize protein language models and supervised machine learning to generate small (n < 100) libraries of D5_AR variants that are subsequently screened for improved neutralization potency. We identify a variant with 5-fold improved neutralization potency, D5_FI, which is the most potent NHR-directed monoclonal antibody characterized to date and exhibits broad neutralization of tier-2 and −3 pseudoviruses as well as replicating R5 and X4 challenge strains. Additionally, our work highlights the ability of protein language models to efficiently identify improved mAb variants from relatively small libraries.

Biopolymers↗

Electrochemical Observation and pH Dependence of All Three Expected Redox Couples in an Extremophilic Bifurcating Electron Transfer Flavoprotein with Fused Subunits

Bifurcating enzymes employ energy from a favorable electron transfer to drive unfavorable transfer of a second electron, thereby generating a more reactive product. They are therefore highly desirable in catalytic systems, for example, to drive challenging reactions such as nitrogen fixation. While most bifurcating enzymes contain air-sensitive metal centers, bifurcating electron transfer flavoproteins (bETFs) employ flavins. However, they have not been successfully deployed on electrodes. Herein, we demonstrate immobilization and expected thermodynamic reactivity of a bETF from a hyperthermophilic archaeon, Sulfolobus acidocaldarius (SaETF). SaETF differs from previously biochemically characterized bETFs in being a single protein, representing a concatenation of the two subunits of known ETFs. However, SaETF retains the chemical properties of heterodimeric bETFs, including possession of two FADs: one that undergoes sequential 1-electron (1e) reductions at high E° and forms an anionic semiquinone, and another that is amenable to lower-E° 2e reduction, including by NADH. We found homologous monomeric ETF genes in archaeal and bacterial genomes, accompanied by genes that also commonly flank heterodimeric ETFs, and SaETF’s sequence conservation is 50% higher with bETFs than with canonical ETFs. Thus, SaETF is best described as a bETF. Our direct electrochemical trials capture reversible redox couples for all three thermodynamically expected redox events. We document electrochemical activity over a range of pH values and reveal a conformational change coupled to proton acquisition that affects the electrochemical activity of the higher-E° FAD. Thus, this well-behaved monomeric bETF opens the door to bioinspired bifurcating devices or bifurcation on a chip.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

The C2 domain augments Ras GTPase-activating protein catalytic activity

Regulation of Ras GTPases by GTPase-activating proteins (GAPs) is essential for their normal signaling. Nine of the ten GAPs for Ras contain a C2 domain immediately proximal to their canonical GAP domain, and in RasGAP (p120GAP, p120RasGAP;RASA1) mutation of this domain is associated with vascular malformations in humans. Here, we show that the C2 domain of RasGAP is required for full catalytic activity toward Ras. Analyses of the RasGAP C2-GAP crystal structure, AlphaFold models, and sequence conservation reveal direct C2 domain interaction with the Ras allosteric lobe. This is achieved by an evolutionarily conserved surface centered around RasGAP residue R707, point mutation of which impairs the catalytic advantage conferred by the C2 domain in vitro. In mice,R707Cmutation phenocopies the vascular and signaling defects resulting from constitutive disruption of theRASA1gene. In SynGAP, mutation of the equivalent conserved C2 domain surface impairs catalytic activity. Our results indicate that the C2 domain is required to achieve full catalytic activity of GAPs for Ras.

Science & Technology - Other Topics↗

Fyn–Saracatinib Complex Structure Reveals an Active State-like Conformation

Fyn is a Src-family tyrosine kinase implicated in synaptic dysfunction and neuroinflammation across multiple neurodegenerative disorders, including Alzheimer’s disease (AD) and Parkinson’s disease (PD). Saracatinib (AZD0530) is a potent Src-family inhibitor that has been explored as a repurposed therapeutic; however, its clinical utility is limited by poor kinase selectivity caused by high sequence conservation within Src-family ATP-binding sites. Here, we combine surface plasmon resonance (SPR) and X-ray crystallography to define saracatinib recognition by the Fyn kinase domain (KD). SPR single-cycle kinetics shows that saracatinib binds the isolated Fyn KD and full-length Fyn with low-nanomolar affinity, whereas dasatinib binds with subnanomolar affinity and markedly slower dissociation. We determined the crystal structure of the Fyn KD-saracatinib complex at 2.22 Å resolution. The kinase adopts an active-like conformation with the DFG motif and αC-helix in the ‘in’ state and a conserved β3 αC Lys-Glu salt bridge. Saracatinib occupies the adenine and ribose pockets, and engages the hinge through direct and water-mediated hydrogen bonding while complementing a hydrophobic back pocket by van der Waals contacts. Comparison with reported saracatinib-bound structures of other kinases suggests that the active-state geometry observed for Fyn creates a pocket not observed in inactive-like complexes, providing a structural handle for designing Fyn-selective inhibitors. Comparison with all saracatinib-bound kinase co-structures currently available in the PDB (ALK2 and PKMYT1) indicates a conserved monodentate hinge binding mode but kinase-dependent αC-helix conformations, providing a structural rationale for designing Fyn-selective analogues.

AZD0530↗

The origin and early evolution of nucleic acid polymerases

The hypothesis that vestiges of the ancestral RNA-dependent RNA polymerase involved in the replication of RNA genomes of Archean cells are present in the eubacterial RNA-polymerase beta-prime subunit and its homologues is discussed. It is shown that, in the DNA-dependent RNA polymerases from three cellular lineages, a very conserved sequence of eight amino acids, also found in a small RNA-binding site previously described for the E. coli polynucleotide phosphorylase and the S1 ribosomal protein, is present. The optimal conditions for the replicase activity of the avian-myeloblastosis-virus reverse transcriptase are presented. The evolutionary significance of the in vitro modifications of substrate and template specificities of RNA polymerases and reverse transcriptases is discussed.

Lazcano, A.↗

Decoding the effects of synonymous variants

Synonymous single nucleotide variants (sSNVs) are common in the human genome but are often overlooked. However, sSNVs can have significant biological impact and may lead to disease. Existing computational methods for evaluating the effect of sSNVs suffer from the lack of gold-standard training/evaluation data and exhibit over-reliance on sequence conservation signals. We developed synVep (synonymous Variant effect predictor), a machine learning-based method that overcomes both of these limitations. Our training data was a combination of variants reported by gnomAD (observed) and those unreported, but possible in the human genome (generated). We used positive-unlabeled learning to purify the generated variant set of any likely unobservable variants. We then trained two sequential extreme gradient boosting models to identify subsets of the remaining variants putatively enriched and depleted in effect. Our method attained 90% precision/recall on a previously unseen set of variants. Furthermore, although synVep does not explicitly use conservation, its scores correlated with evolutionary distances between orthologs in cross-species variation analysis. synVep was also able to differentiate pathogenic vs. benign variants, as well as splice-site disrupting variants (SDV) vs. non-SDVs. Thus, synVep provides an important improvement in annotation of sSNVs, allowing users to focus on variants that most likely harbor effects.

Zishuo Zeng↗

Two deeply conserved non-coding sequences control PLETHORA1/2 expression and coordinate embryo and root development

Conserved non-coding sequences (CNSs) are integral elements of transcriptional regulation. Transcriptional tuning of PLETHORA (PLT) genes that encode master regulators of plant development is vital for embryogenesis and meristematic function. However, how the expression of PLT genes is modulated through CNSs remains unclear. Through motif-based mining of upstream sequences in 120 angiosperm genomes, we identified 21 conserved and lineage-specific CNSs, two of which are unusually long, similar, and colinear within eudicots. Using Arabidopsis thaliana, we demonstrate that these two deeply conserved elements, which we named BOX1 and BOX2, control PLT1 and PLT2 expression. CRISPR mutants within these elements specifically reduced PLT expression levels, and reporter lines revealed that deletion of either or both BOXes altered and/or abrogated the PLT2 expression pattern in the root tip, affecting the ability to rescue the plt1 plt2 double mutant. We further show that the influence of these elements on expression patterns is already exerted during embryogenesis and functional in the context of the early embryo. Finally, we reveal the existence of a BOX-mediated autoregulatory feedback loop that, in large part, explains CNS influence on expression patterns. We thus uncover a transcriptional mechanism by which genes encoding master regulators of embryo and root meristem development are regulated.

PLETHORA↗

Co-conservation of rRNA tetraloop sequences and helix length suggests involvement of the tetraloops in higher-order interactions

Terminal loops containing four nucleotides (tetraloops) are common in structural RNAs, and they frequently conform to one of three sequence motifs, GNRA, UNCG, or CUUG. Here we compare available sequences and secondary structures for rRNAs from bacteria, and we show that helices capped by phylogenetically conserved GNRA loops display a strong tendency to be of conserved length. The simplest interpretation of this correlation is that the conserved GNRA loops are involved in higher-order interactions, intramolecular or intermolecular, resulting in a selective pressure for maintaining the lengths of these helices. A small number of conserved UNCG loops were also found to be associated with conserved length helices, consistent with the possibility that this type of tetraloop also takes part in higher-order interactions.

Non-NASA Center↗

Conservative Patch Algorithm and Mesh Sequencing for PAB3D

A mesh-sequencing algorithm and a conservative patched-grid-interface algorithm (hereafter Patch Algorithm ) have been incorporated into the PAB3D code, which is a computer program that solves the Navier-Stokes equations for the simulation of subsonic, transonic, or supersonic flows surrounding an aircraft or other complex aerodynamic shapes. These algorithms are efficient, flexible, and have added tremendously to the capabilities of PAB3D. The mesh-sequencing algorithm makes it possible to perform preliminary computations using only a fraction of the grid cells (provided the original cell count is divisible by an integer) along any grid coordinate axis, independently of the other axes. The patch algorithm addresses another critical need in multi-block grid situation where the cell faces of adjacent grid blocks may not coincide, leading to errors in calculating fluxes of conserved physical quantities across interfaces between the blocks. The patch algorithm, based on the Stokes integral formulation of the applicable conservation laws, effectively matches each of the interfacial cells on one side of the block interface to the corresponding fractional cell area pieces on the other side. This approach is comprehensive and unified such that all interface topology is automatically processed without user intervention. This algorithm is implemented in a preprocessing code that creates a cell-by-cell database that will maintain flux conservation at any level of full or reduced grid density as the user may choose by way of the mesh-sequencing algorithm. These two algorithms have enhanced the numerical accuracy of the code, reduced the time and effort for grid preprocessing, and provided users with the flexibility of performing computations at any desired full or reduced grid resolution to suit their specific computational requirements.

Pao, S. P.↗

The landscape of regulatory element evolution in a C4 perennial grass

Gene regulatory evolution is a well-known source of phenotypic diversity and adaptive evolution. Although cis-regulatory elements (CREs) play a vital role in gene expression evolution, the molecular evolution of CREs remains mostly unknown due to the difficulty in identifying and characterizing these functional elements. Comparative genomic analyses of noncoding DNA can be leveraged to identify conserved noncoding sequences (CNS), many of which may harbor functional CREs conserved by purifying selection. However, purely computational inference of CREs from putative CNS can be erroneous due to the complex genomic architecture in plants. One promising experimental approach to identify CREs is by profiling accessible chromatin regions (ACRs) that are often associated with the location of CREs. In this study, we use comparative genomics along with the profiling of ACRs to study the molecular evolution of putative functional noncoding regulatory regions in Panicoid grasses. We identified sets of CNS that varied in relationship to the degree of evolutionary divergence among the studied taxa, including identifying core-Panicoid-CNS. We augmented this analysis by profiling ACRs in Panicum hallii ecotypes using ATAC-seq. ACRs had low SNP density at the summit, harbored a high frequency of core-Panicoid-CNS, and were enriched with expression QTL. These data help to annotate the P. hallii genome for putative functional elements and suggest that a large proportion of these ACRs are evolving under purifying selection. Turnover in CNS and ACR between ecotypes of P. hallii identifies a small set of putatively divergent CREs that may underlie differences in gene regulation between genotypes from inland and coastal habitats. In summary, we profiled ACRs in Panicoid grasses and integrated this data with our putative CNS prediction framework, which provides unique insight into patterns of polymorphism and divergence in CREs in C4 perennial grasses.

59 BASIC BIOLOGICAL SCIENCES↗

Structurally complex and highly active RNA ligases derived from random RNA sequences

Seven families of RNA ligases, previously isolated from random RNA sequences, fall into three classes on the basis of secondary structure and regiospecificity of ligation. Two of the three classes of ribozymes have been engineered to act as true enzymes, catalyzing the multiple-turnover transformation of substrates into products. The most complex of these ribozymes has a minimal catalytic domain of 93 nucleotides. An optimized version of this ribozyme has a kcat exceeding one per second, a value far greater than that of most natural RNA catalysts and approaching that of comparable protein enzymes. The fact that such a large and complex ligase emerged from a very limited sampling of sequence space implies the existence of a large number of distinct RNA structures of equivalent complexity and activity.

Non-NASA Center↗

Structural models of the MscL gating mechanism

Three-dimensional structural models of the mechanosensitive channel of large conductance, MscL, from the bacteria Mycobacterium tuberculosis and Escherichia coli were developed for closed, intermediate, and open conformations. The modeling began with the crystal structure of M. tuberculosis MscL, a homopentamer with two transmembrane alpha-helices, M1 and M2, per subunit. The first 12 N-terminal residues, not resolved in the crystal structure, were modeled as an amphipathic alpha-helix, called S1. A bundle of five parallel S1 helices are postulated to form a cytoplasmic gate. As membrane tension induces expansion, the tilts of M1 and M2 are postulated to increase as they move away from the axis of the pore. Substantial expansion is postulated to occur before the increased stress in the S1 to M1 linkers pulls the S1 bundle apart. During the opening transition, the S1 helices and C-terminus amphipathic alpha-helices, S3, are postulated to dock parallel to the membrane surface on the perimeter of the complex. The proposed gating mechanism reveals critical spatial relationships between the expandable transmembrane barrel formed by M1 and M2, the gate formed by S1 helices, and "strings" that link S1s to M1s. These models are consistent with numerous experimental results and modeling criteria.

NASA Discipline Cell Biology↗

High phenotypic and genotypic plasticity among strains of the mushroom-forming fungus Schizophyllum commune

Schizophyllum commune is a mushroom-forming fungus notable for its distinctive fruiting bodies with split gills. It is used as a model organism to study mushroom development, lignocellulose degradation and mating type loci. It is a hypervariable species with considerable genetic and phenotypic diversity between the strains. In this study, we systematically phenotyped 16 dikaryotic strains for aspects of mushroom development and 18 monokaryotic strains for lignocellulose degradation. There was considerable heterogeneity among the strains regarding these phenotypes. The majority of the strains developed mushrooms with varying morphologies, although some strains only grew vegetatively under the tested conditions. Growth on various carbon sources showed strain-specific profiles. The genomes of seven monokaryotic strains were sequenced and analyzed together with six previously published genome sequences. Moreover, the related species Schizophyllum fasciatum was sequenced. Although there was considerable genetic variation between the genome assemblies, the genes related to mushroom formation and lignocellulose degradation were well conserved. These sequenced genomes, in combination with the high phenotypic diversity, will provide a solid basis for functional genomics analyses of the strains of S. commune.

59 BASIC BIOLOGICAL SCIENCES↗

A lignin-specific peroxidase in tobacco whose antisense suppression leads to vascular tissue modification

A tobacco peroxidase isoenzyme (TP60) was down-regulated in tobacco using an antisense strategy, this affording transformants with lignin reductions of up to 40-50% of wild type (control) plants. Significantly, both guaiacyl and syringyl levels decreased in essentially a linear manner with the reductions in lignin amounts, as determined by both thioacidolysis and nitrobenzene oxidative analyses. These data provisionally suggest that a feedback mechanism is operative in lignifying cells, which prevents build-up of monolignols should oxidative capacity for their subsequent metabolism be reduced. Prior to this study, the only known rate-limiting processes in the monolignol/lignin pathways involved that of Phe supply and the relative activities of cinnamate-4-hydroxylase/p-coumarate-3-hydroxylase, respectively. These transformants thus provide an additional experimental means in which to further dissect and delineate the factors involved in monolignol targeting to precise regions in the cell wall, and of subsequent lignin assembly. Interestingly, the lignin down-regulated tobacco phenotypes displayed no readily observable differences in overall growth and development profiles, although the vascular apparatus was modified.

NASA Program Fundamental Space Biology↗

ALS mutations disrupt self-association between the ubiquilin STI1 hydrophobic groove and internal placeholder sequences

Ubiquilins are molecular chaperones that play multifaceted roles in proteostasis, with point mutations in UBQLN2 leading to altered phase-separation properties and amyotrophic lateral sclerosis (ALS). Our mechanistic understanding of this essential process has been hindered by a lack of structural information on the STI1 domain, which is essential for ubiquilin chaperone activity and phase separation. Here, we present the first crystal structure of a ubiquilin-family STI1 domain bound to a transmembrane domain (TMD), and show that ALS mutations disrupt the STI1-TMD interaction. We further demonstrate that ubiquilins contain multiple conserved internal sequences that bind to the STI1 domain, including the PXX-repeat region that is a hotspot for ALS mutations. We propose that these placeholder sequences prevent solvent exposure of the STI1 hydrophobic groove and contribute to the multivalency that drives ubiquilin phase-separation. Together, this work provides a new paradigm for understanding how STI1 domains modulate ubiquilin chaperone activity and phase separation, and offers insights into the molecular basis of ALS pathogenesis.

Onwunma, Joan [Univ. of Toledo, OH (United States)↗