Search NASA⌕ Search

SEARCH · Search NASA

Results for “Conserved Sequence”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

Conservation of Pax gene expression in ectodermal placodes of the lamprey

Ectodermal placodes contribute to the cranial ganglia and sense organs of the head and, together with neural crest cells, represent defining features of the vertebrate embryo. The identity of different placodes appears to be specified in part by the expression of different Pax genes, with Pax-3/7 class genes being expressed in the trigeminal placode of mice, chick, frogs and fish, and Pax-2/5/8 class genes expressed in the otic placode. Here, we present the cloning and expression pattern of lamprey Pax-7 and Pax-2, which mark the trigeminal and otic placodes, respectively, as well as other structures characteristic of vertebrate Pax genes. These results suggest conservation of Pax genes and placodal structures in basal and derived vertebrates.

NASA Discipline Evolutionary Biology↗

A genomic perspective on fungal diversity and evolution

Originating from aquatic unicellular ancestors, over the course of ~1 billion years, the fungi have evolved to occupy nearly all aerobic environments on the planet, diversified into millions of different ‘species’ and have developed complex multicellular structures. Their relatively small, simple genomes have facilitated massive-scale sequencing and allowed us to explore genome evolution across an ancient eukaryotic kingdom. With thousands of genomes from diverse lineages now available, this Review will discuss insights into fungal biology and evolution gleaned with genomics and other multi-omics approaches. Using published genomes available through GenBank and the Joint Genome Institute’s MycoCosm platform, we generated kingdom-wide phylogenies and used them to highlight how fungal genomes have changed over time. With this phylogeny as a guide, we also discuss major evolutionary transitions that occurred across the fungal kingdom. Although progress has been made, these efforts are hampered by biases in genome representation and limited characterization of gene functions. Here, in this study, we discuss these challenges and possible future directions to address them, including initiatives to characterize conserved genes of unknown function and scale up sequencing towards 10,000 annotated fungal genomes.

Mondo, Stephen J. [USDOE Joint Genome Institute (J↗

Mitochondrial gene arrangement of the horseshoe crab Limulus polyphemus L.: conservation of major features among arthropod classes

Numerous complete mitochondrial DNA sequences have been determined for species within two arthropod groups, insects and crustaceans, but there are none for a third, the chelicerates. Most mitochondrial gene arrangements reported for crustaceans and insect species are identical or nearly identical to that of Drosophila yakuba. Sequences across 36 of the gene boundaries in the mitochondrial DNA (mtDNA) of a representative chelicerate. Limulus polyphemus L., also reveal an arrangement like that of Drosophila yakuba. Only the position of the tRNA(LEU)(UUR) gene differs; in Limulus it is between the genes for tRNA(LEU)(CUN) and ND1. This positioning is also found in onychophorans, mollusks, and annelids, but not in insects and crustaceans, and indicates that tRNA(LEU)(CUN)-tRNA(LEU)(UUR)-ND1 was the ancestral gene arrangement for these groups, as suggested earlier. There are no differences in the relative arrangements of protein-coding and ribosomal RNA genes between Limulus and Drosophila, and none have been observed within arthropods. The high degree of similarity of mitochondrial gene arrangements within arthropods is striking, since some taxa last shared a common ancestor before the Cambrian, and contrasts with the extensive mtDNA rearrangements occasionally observed within some other metazoan phyla (e.g., mollusks and nematodes).

Non-NASA Center↗

Conservation of Endo16 expression in sea urchins despite evolutionary divergence in both cis and trans-acting components of transcriptional regulation

Evolutionary changes in transcriptional regulation undoubtedly play an important role in creating morphological diversity. However, there is little information about the evolutionary dynamics of cis-regulatory sequences. This study examines the functional consequence of evolutionary changes in the Endo16 promoter of sea urchins. The Endo16 gene encodes a large extracellular protein that is expressed in the endoderm and may play a role in cell adhesion. Its promoter has been characterized in exceptional detail in the purple sea urchin, Strongylocentrotus purpuratus. We have characterized the structure and function of the Endo16 promoter from a second sea urchin species, Lytechinus variegatus. The Endo16 promoter sequences have evolved in a strongly mosaic manner since these species diverged approximately 35 million years ago: the most proximal region (module A) is conserved, but the remaining modules (B-G) are unalignable. Despite extensive divergence in promoter sequences, the pattern of Endo16 transcription is largely conserved during embryonic and larval development. Transient expression assays demonstrate that 2.2 kb of upstream sequence in either species is sufficient to drive GFP reporter expression that correctly mimics this pattern of Endo16 transcription. Reciprocal cross-species transient expression assays imply that changes have also evolved in the set of transcription factors that interact with the Endo16 promoter. Taken together, these results suggest that stabilizing selection on the transcriptional output may have operated to maintain a similar pattern of Endo16 expression in S. purpuratus and L. variegatus, despite dramatic divergence in promoter sequence and mechanisms of transcriptional regulation.

NASA Discipline Evolutionary Biology↗

A pressure correction method for the calculation of compressible chemical reacting flows

A recently developed noniterative method for the solution of the transient fluid flow equations at all speed is extended to handle chemical reacting flows. The species conservation equations are loosely coupled into the predictor/multicorrector sequence of the solution procedure. A split-operator method separates the chemical kinetics terms from the fluid-dynamical terms, as well as an implicit differencing method enhance the numerical stability. The method was applied for turbulent diffusion flame calculations and for the analyses of high pressure, axisymmetric turbulent hypersonic nozzle flows. The diffusion flame results were compared with a similar pressure method for fast chemistry integration scheme without operator-splitting. Simulations of the nozzle flow indicated that the nonideal intermolecular effects must be included in the analysis and design of high pressure hypersonic nozzle.

Chen, Z. J.↗

Surfactant-like peptide gels are based on cross-β amyloid fibrils

Surfactant-like peptides, in which hydrophilic and hydrophobic residues are encoded within different domains in the peptide sequence, undergo facile self-assembly in aqueous solution to form supramolecular hydrogels. These peptides have been explored extensively as substrates for the creation of functional materials since a wide variety of amphipathic sequences can be prepared from commonly available amino acid precursors. The self-assembly behavior of surfactant-like peptides has been compared to that observed for small molecule amphiphiles in which nanoscale phase separation of the hydrophobic domains drives the self-assembly of supramolecular structures. Here, we investigate the relationship between sequence and supramolecular structure for a pair of bola-amphiphilic peptides, Ac-KLIIIK-NH 2 (L2) and Ac-KIIILK-NH 2 (L5). Despite similar length, composition, and polar sequence pattern, L2 and L5 form morphologically distinct assemblies, nanosheets and nanotubes, respectively. Cryo-EM helical reconstruction was employed to determine the structure of the L5 nanotube at near-atomic resolution. Rather than displaying self-assembly behavior analogous to conventional amphiphiles, the packing arrangement of peptides in the L5 nanotube displayed steric zipper interfaces that resembled those observed in the structures of β-amyloid fibrils. Like amyloids, the supramolecular structures of the L2 and L5 assemblies were sensitive to conservative amino acid substitutions within an otherwise identical amphipathic sequence pattern. This study highlights the need to better understand the relationship between sequence and supramolecular structure to facilitate the development of functional peptide-based materials for biomaterials applications.

Das, Abhinaba [Emory University, Atlanta, GA (Unit↗

On Betz' Method for Rollup of Vortex Sheets

A method introduced by Betz in 1932 relates the vortex sheet shed by one side of a lifting wing to the rolled-up vortex far downstream. The Betz rollup theory came into active use during the 1970's when the hazard posed by vortex wakes of subsonic transport aircraft became of concern at airports. Even though the method involves several simplifying assumptions, it provides insight into the rollup process and has proven to be a useful and reliable tool for the analysis and diagnosis of lift-generated wakes. As part of an ongoing effort to develop improved guidelines for the rollup process of complex vortex wakes, the research described is directed at the study of the two vortex invariants not used in the Betz formulation, and to find out if they can be used to determine the sequence of rollup or the center of vortices where rollup begins. It is found that the two unused vortex invariants are both conserved during rollup, but that they do not yield any information or guidance on rollup sequence or on the location of vortex centers. It is also found that the invariant for energy can be used to determine the rolled-up structure of vortices but the structure differs negligibly from that predicted by use of the invariant for the second moment of circulation.

Rossow, Vernon J.↗

Naturally ornate RNA-only complexes revealed by cryo-EM

The structures of natural RNAs remain poorly characterized and may hold numerous surprises. Here we report three-dimensional structures of three large ornate bacterial RNAs using cryo-electron microscopy (cryo-EM). GOLLD (Giant, Ornate, Lake- and Lactobacillales-Derived), ROOL (Rumen-Originating, Ornate, Large) and OLE (Ornate Large Extremophilic) RNAs form homo-oligomeric complexes whose stoichiometries are retained at lower concentrations than measured in cells. OLE RNA forms a dimeric complex with long co-axial pipes spanning two monomers. Both GOLLD and ROOL form distinct RNA-only multimeric nanocages with diameters larger than the ribosome, each empty except for a disordered loop. Extensive intramolecular and intermolecular A-minor interactions, kissing loops, an unusual A–A helix and other interactions stabilize the three complexes. Sequence covariation analysis of these large RNAs reveals evolutionary conservation of intermolecular interactions, supporting the biological importance of large, ornate RNA quaternary structures that can assemble without any involvement of proteins.

59 BASIC BIOLOGICAL SCIENCES↗

Size and Structure of the Sequence Space of Repeat Proteins

The coding space of protein sequences is shaped by evolutionary constraints set by requirements of function and stability. We show that the coding space of a given protein family— the total number of sequences in that family—can be estimated using models of maximum entropy trained on multiple sequence alignments of naturally occurring amino acid sequences. We analyzed and calculated the size of three abundant repeat proteins families, whose members are large proteins made of many repetitions of conserved portions of *30 amino acids. While amino acid conservation at each position of the alignment explains most of the reduction of diversity relative to completely random sequences, we found that correlations between amino acid usage at different positions significantly impact that diversity. We quantified the impact of different types of correlations, functional and evolutionary, on sequence diversity. Analysis of the detailed structure of the coding space of the families revealed a rugged landscape, with many local energy minima of varying sizes with a hierarchical structure, reminiscent of frustrated energy landscapes of spin glass in physics. This clustered structure indicates a multiplicity of subtypes within each family and suggests new strategies for protein design.

Jacopo Marchi↗

Cholesterol-dependent enzyme activity of human TSPO1

The amino acid sequence of the tryptophan-rich sensory proteins (TSPO) is substantially conserved throughout all kingdoms of life. Human mitochondrial TSPO1 (HsTSPO1) binds to porphyrins and steroids, although its interactions with these molecules remains unknown.HsTSPO1 is associated with numerous physiological and pathological disorders, but the underlying molecular mechanisms are unknown. Here, we disclose the finding of human mitochondrial TSPO as a cholesterol-dependent protoporphyrin IX oxygenase. The results of our biochemical characterization are consistent with structural data and evolutionary analysis. The dependence ofHsTSPO1 activity on cholesterol may be the result of the coevolution of this membrane protein with the membrane system. Our study provides a molecular foundation for comprehending the various roles played by mitochondrial TSPO in normal physiological and pathological situations.

Science & Technology - Other Topics↗

Borg extrachromosomal elements of methane-oxidizing archaea have conserved and expressed genetic repertoires

Borgs are huge extrachromosomal elements (ECE) of anaerobic methane-consuming “Candidatus Methanoperedens” archaea. Here, we used nanopore sequencing to validate published complete genomes curated from short reads and to reconstruct new genomes. 13 complete and four near-complete linear genomes share 40 genes that define a largely syntenous genome backbone. We use these conserved genes to identify new Borgs from peatland soil and to delineate Borg phylogeny, revealing two major clades. Remarkably, Borg genes encoding nanowire-like electron-transferring cytochromes and cell surface proteins are more highly expressed than those of host Methanoperedens, indicating that Borgs augment the Methanoperedens activity in situ. We reconstructed the first complete 4.00 Mbp genome for a Methanoperedens that is inferred to be a Borg host and predicted its methylation motifs, which differ from pervasive TC and CC methylation motifs of the Borgs. Thus, methylation may enable Methanoperedens to distinguish their genomes from those of Borgs. Very high Borg to Methanoperedens ratios and structural predictions suggest that Borgs may be capable of encapsulation. The findings clearly define Borgs as a distinct class of ECE with shared genomic signatures, establish their diversification from a common ancestor with genetic inheritance, and raise the possibility of periodic existence outside of host cells.

59 BASIC BIOLOGICAL SCIENCES↗

Hemichordates and the Origin of Chordates

At the start of the period of the NASA grant three years ago, we had no information on the organization and development of the body axis of the hemichordate, Saccoglossus kowalevskii. Now we have substantial findings about the anteroposterior axis and dorsoventral axis, and based on this information, we have new insights about the origin of chordates from ancestral deuterostomes. We found ways to obtain and preserve large numbers of embryos and hatched juveniles. We can now collect about 40,000 embryos in the month of September, the time of S. kowalevskii spawning at Woods Hole. Excellent cDNA libraries were prepared from three developmental stages. From these libraries, we directly isolated about 30 gene ortholog sequences by screening and pcr techniques, all of these sequences of interest in the inquiry about the animal's organization and development. We also performed a mid-sized EST project (60,000 randomly picked clones, many of these arrayed). About half of these have been analyzed so far by blastx and are suitable for direct use of clones. We have obtained about 50 interesting sequences from this set. The rest still await analysis. Thus, at this time we have isolated orthologs of 80 genes that are known to be expressed in chordates in conserved domains and known to have interesting roles in chordate organization and development. The orthology of the S. kowalevskii sequences has been verified by neighbor joining and parsimony methods, with bootstrap estimates of validity. The S. kowalevskii sequences cluster with other deuterostome sequences, namely, other hemichordates, echinoderms, ascidians, amphioxus, or vertebrates, depending on what sequences are available in the database for comparison. We have used these sequences to do high quality in situ hybridization on S. kowalevskii embryos, and the results can be divided into three sections-those concerning the anteroposterior axis of S. kowalevskii in comparison to the same axis of chordates, those concerning the dorsoventral axis of S. kowalevskii in comparison to the same axis of chordates, and those concerning the signals and transcription factors found in the endoderm, of S. kowalevskii compared to the signals and transcription factors in the endo-mesodermal cells of Spemann's organizer of chordates.

Gerhart, John↗

Genes encoding calmodulin-binding proteins in the Arabidopsis genome

Analysis of the recently completed Arabidopsis genome sequence indicates that approximately 31% of the predicted genes could not be assigned to functional categories, as they do not show any sequence similarity with proteins of known function from other organisms. Calmodulin (CaM), a ubiquitous and multifunctional Ca(2+) sensor, interacts with a wide variety of cellular proteins and modulates their activity/function in regulating diverse cellular processes. However, the primary amino acid sequence of the CaM-binding domain in different CaM-binding proteins (CBPs) is not conserved. One way to identify most of the CBPs in the Arabidopsis genome is by protein-protein interaction-based screening of expression libraries with CaM. Here, using a mixture of radiolabeled CaM isoforms from Arabidopsis, we screened several expression libraries prepared from flower meristem, seedlings, or tissues treated with hormones, an elicitor, or a pathogen. Sequence analysis of 77 positive clones that interact with CaM in a Ca(2+)-dependent manner revealed 20 CBPs, including 14 previously unknown CBPs. In addition, by searching the Arabidopsis genome sequence with the newly identified and known plant or animal CBPs, we identified a total of 27 CBPs. Among these, 16 CBPs are represented by families with 2-20 members in each family. Gene expression analysis revealed that CBPs and CBP paralogs are expressed differentially. Our data suggest that Arabidopsis has a large number of CBPs including several plant-specific ones. Although CaM is highly conserved between plants and animals, only a few CBPs are common to both plants and animals. Analysis of Arabidopsis CBPs revealed the presence of a variety of interesting domains. Our analyses identified several hypothetical proteins in the Arabidopsis genome as CaM targets, suggesting their involvement in Ca(2+)-mediated signaling networks.

NASA Discipline Plant Biology↗

The reference genome for the northeastern Pacific bull kelp, Nereocystis luetkeana

Bull kelp, Nereocystis luetkeana, is a northeastern Pacific kelp with broad distribution from Alaska to central California. Its population declines have caused severe concerns in northern California, the Salish Sea in Washington, and recently in some populations in Oregon. Despite bull kelp's accumulated ecological and physiological studies, an assembled and annotated genomic reference was still unavailable. Here, we report the complete and annotated genome of Nereocystis luetkeana, produced by the California Conservation Genomics Project (CCGP), which aims to reveal genomic diversity patterns across California by sequencing the complete genomes of approximately 150 carefully selected species. The genome was assembled into 1562 scaffolds with 449.82 Mb, 80x of coverage and 22 952 gene models. BUSCO assembly showed a completeness score of 72% for the stramenopiles gene set. The mitochondria and chloroplast genome sequences have 37 Kb and 131 Mb, respectively. The orthology analysis between 10 Phaeophycean genomes showed 1065 expanded and 286 unique orthogroups for this species. Pairwise comparisons showed 542 orthogroups present only in N. luetkeana and M. pyrifera, another large-body kelp. The enrichment analysis of these orthogroups showed important functions related to central metabolism and signaling due to ATPases enrichment in these two species. This genome assembly will provide an essential resource for the ecology, evolution, conservation, and breeding of bull kelp.

California Conservation Genomics Project—CCGP↗

Comparative proteomics of a versatile, marine, iron-oxidizing chemolithoautotroph

This study conducted a comparative proteomic analysis to identify potential genetic markers for the biological function of chemolithoautotrophic iron oxidation in the marine bacterium Ghiorsea bivora. To date, this is the only characterized species in the class Zetaproteobacteria that is not an obligate iron-oxidizer, providing a unique opportunity to investigate differential protein expression to identify key genes involved in iron-oxidation at circumneutral pH. Over 1000 proteins were identified under both iron- and hydrogen-oxidizing conditions, with differentially expressed proteins found in both treatments. Notably, a gene cluster upregulated during iron oxidation was identified. This cluster contains genes encoding for cytochromes that share sequence similarity with the known iron-oxidase, Cyc2. Interestingly, these cytochromes, conserved in both Bacteria and Archaea, do not exhibit the typical β-barrel structure of Cyc2. This cluster potentially encodes a biological nanowire-like transmembrane complex containing multiple redox proteins spanning the inner membrane, periplasm, outer membrane, and extracellular space. The upregulation of key genes associated with this complex during iron-oxidizing conditions was confirmed by quantitative reverse transcription-PCR. These findings were further supported by electromicrobiological methods, which demonstrated negative current production by G. bivora in a three-electrode system poised at a cathodic potential. This research provides significant insights into the biological function of chemolithoautotrophic iron oxidation.

59 BASIC BIOLOGICAL SCIENCES↗

Conformational heterogeneity in the dGsw purine riboswitch: role of Mg²⁺ and 2’-dG in aptamer folding

Recent advancements in RNA structural biology have focused on unraveling the complexities of non-coding mRNA elements like riboswitches. These cis-acting regulatory regions undergo structural changes in response to specific cellular metabolites, leading to up or downregulation of downstream genes. The purine riboswitch family regulates many prokaryotic genes involved in purine degradation and biosynthesis. They feature an aptamer domain organized around a 3-way helical junction, where ligand encapsulation occurs at the junctional core. In our study, we chemically probed the aptamer domain of the 2’-dG-sensing purine riboswitch from Mesoplasma florum (dGsw) under various solution conditions to understand how Mg²⁺ and 2’-dG influence riboswitch folding. Here, we find that efficient 2’-dG binding strongly depends on Mg²⁺, indicating that Mg²⁺ is essential for priming dGsw for ligand interactions. We identified a previously undescribed sequence in the 5’ tail of dGsw that is complementary to a conserved helix. The inclusion of this region in a construct led to intramolecular competition between the alternate helix, Palt, and P1. Mutational analysis confirmed that 5’ flanking end of the aptamer domain forms an alternate helix in the absence of ligand. Molecular dynamics simulations revealed that this alternative conformation is stable. This helix may, therefore, facilitate the formation of an anti-terminator helix by opening the 3-way junction surrounding the 2’-dG binding site. Our study further establishes the importance of a closed terminal P1 helix conformation for metabolite binding and suggests that the delicate interplay between P1 and Palt may fine-tune downstream gene regulation. These insights offer a new perspective on riboswitch structure and enhance our understanding of the role that a conformational ensemble plays in riboswitch activity and regulation.

Biochemistry & Molecular Biology↗

A conserved viral RNA fold enables nuclease resistance across kingdoms of life

Abstract Viral exoribonuclease-resistant RNA (xrRNA) structures block cellular nucleases to produce subgenomic viral RNAs during infection. High sequence variability among xrRNAs from distantly related viruses raises questions about the shared molecular features that enable these RNAs to withstand the strong unwinding forces of exoribonucleases. Here, we present the first structure of a plant-virus xrRNA in its active conformation and uncover universal principles of xrRNA folding. Comparison with the structure of a human-pathogenic flavivirus xrRNA reveals that both share a core structural motif—a protective ring encircling the RNA’s 5′ end—despite lacking sequence similarity. Disrupting this core motif through targeted mutagenesis eliminates exoribonuclease-resistance and attenuates viral infection. We identify hundreds of related structures across multiple virus families, supporting the conservation of this mechanism. Our study demonstrates how distantly related RNA viruses have converged on a common structural strategy to inhibit cellular nucleases, with a universal ring topology as the defining feature of viral xrRNAs.

Biochemistry & Molecular Biology↗

Calmodulin activation of an endoplasmic reticulum-located calcium pump involves an interaction with the N-terminal autoinhibitory domain

To investigate how calmodulin regulates a unique subfamily of Ca(2+) pumps found in plants, we examined the kinetic properties of isoform ACA2 identified in Arabidopsis. A recombinant ACA2 was expressed in a yeast K616 mutant deficient in two endogenous Ca(2+) pumps. Orthovanadate-sensitive (45)Ca(2+) transport into vesicles isolated from transformants demonstrated that ACA2 is a Ca(2+) pump. Ca(2+) pumping by the full-length protein (ACA2-1) was 4- to 10-fold lower than that of the N-terminal truncated ACA2-2 (Delta2-80), indicating that the N-terminal domain normally acts to inhibit the pump. An inhibitory sequence (IC(50) = 4 microM) was localized to a region within valine-20 to leucine-44, because a peptide corresponding to this sequence lowered the V(max) and increased the K(m) for Ca(2+) of the constitutively active ACA2-2 to values comparable to the full-length pump. The peptide also blocked the activity (IC(50) = 7 microM) of a Ca(2+) pump (AtECA1) belonging to a second family of Ca(2+) pumps. This inhibitory sequence appears to overlap with a calmodulin-binding site in ACA2, previously mapped between aspartate-19 and arginine-36 (J.F. Harper, B. Hong, I. Hwang, H.Q. Guo, R. Stoddard, J.F. Huang, M.G. Palmgren, H. Sze inverted question mark1998 J Biol Chem 273: 1099-1106). These results support a model in which the pump is kept "unactivated" by an intramolecular interaction between an autoinhibitory sequence located between residues 20 and 44 and a site in the Ca(2+) pump core that is highly conserved between different Ca(2+) pump families. Results further support a model in which activation occurs as a result of Ca(2+)-induced binding of calmodulin to a site overlapping or immediately adjacent to the autoinhibitory sequence.

NASA Program Fundamental Space Biology↗