Search NASA⌕ Search

SEARCH · Search NASA

Results for “Sequence”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

Plant sulfate transporter protein sequences for phylogenetic analysis

Sulfur is an essential macronutrient that supports plant growth, development, and responses to environmental stress. Sulfate is the predominant inorganic form of sulfur in soils, and its uptake by roots and translocation to shoots are facilitated by the sulfate transporter (SULTR) family of proteins. Although the first plant SULTR gene was identified nearly three decades ago, several subfamily members, particularly those in the expansive and angiosperm-specific SULTR3 group, remain poorly characterized. To support comprehensive phylogenetic and sequence-based analyses, we compiled a curated dataset of 262 SULTR protein sequences from 22 plant species spanning the evolutionary breadth of land plants. This collection includes representatives from two basal lineages, two early-divergent angiosperms, six monocots, and ten dicots. All sequences were extracted from genome assemblies available in Phytozome v13 (Joint Genome Institute) and manually curated, with cross-referencing to additional databases such as NCBI when needed. This dataset provides a valuable resource for reconstructing the evolutionary history of the SULTR family, with particular emphasis on the diversification of SULTR3 transporters in flowering plants. This resource may also support functional annotation, comparative genomics, and structural modeling of sulfate transport proteins.

CBI↗

Genomic-based biosurveillance for avian influenza: whole genome sequencing from wild mallards sampled during autumn migration in 2022–2023 reveals a high co-infection rate on migration stopover site in Georgia

The Caucasus region, including Georgia, is an important intersection for migratory waterbirds, offering potential for avian influenza virus (AIV) transmission between populations from different geographic areas. In 2022 and 2023, wild ducks were sampled during autumn migration events in Georgia to study the genetic relationships and molecular characteristics of influenza strains. Sequencing and phylogenetic analysis were used to compare the sampled strains to reference sequences from Africa, Asia, and Europe, allowing assessment of genetic relationships and virus transmission between migratory birds. Protein language modeling identified potential co-infections. Of 225 duck samples, 128 tested positive for the influenza M gene. 55 influenza-positive samples underwent whole-genome sequencing, revealing significant diversity. Analysis of the hemagglutinin (HA) segment showed notable differences among subtypes. Most samples were H6N1 and H6N6, but co-infections with combinations like H6H3, N8N1, N6H9, N2N6, and H9H6/N1N2 were also identified. These findings demonstrate the high variability of influenza viruses in migratory waterbirds in Georgia, including a notable rate of co-infections. Some samples exhibited uncommon genetic characteristics compared to other strains from the same year, suggesting Georgia’s role as a mixing vessel for influenza viruses. This facilitates reassortment during co-infections and contributes to the genetic diversity observed across flyways.

59 BASIC BIOLOGICAL SCIENCES↗

Two deeply conserved non-coding sequences control PLETHORA1/2 expression and coordinate embryo and root development

Conserved non-coding sequences (CNSs) are integral elements of transcriptional regulation. Transcriptional tuning of PLETHORA (PLT) genes that encode master regulators of plant development is vital for embryogenesis and meristematic function. However, how the expression of PLT genes is modulated through CNSs remains unclear. Through motif-based mining of upstream sequences in 120 angiosperm genomes, we identified 21 conserved and lineage-specific CNSs, two of which are unusually long, similar, and colinear within eudicots. Using Arabidopsis thaliana, we demonstrate that these two deeply conserved elements, which we named BOX1 and BOX2, control PLT1 and PLT2 expression. CRISPR mutants within these elements specifically reduced PLT expression levels, and reporter lines revealed that deletion of either or both BOXes altered and/or abrogated the PLT2 expression pattern in the root tip, affecting the ability to rescue the plt1 plt2 double mutant. We further show that the influence of these elements on expression patterns is already exerted during embryogenesis and functional in the context of the early embryo. Finally, we reveal the existence of a BOX-mediated autoregulatory feedback loop that, in large part, explains CNS influence on expression patterns. We thus uncover a transcriptional mechanism by which genes encoding master regulators of embryo and root meristem development are regulated.

PLETHORA↗

One-Pot Self-Assembly of Sequence-Controlled Mesoporous Heterostructures via Structure-Directing Agents

Multimaterial heterostructures have led to characteristics surpassing the individual components. Nature controls the architecture and placement of multiple materials through biomineralization of nanoparticles (NPs); however, synthetic heterostructure formation remains limited and generally departs from the elegance of self-assembly. Here, in this study, a class of block polymer structure-directing agents (SDAs) are developed containing repeat units capable of persistent (covalent) NP interactions that enable the direct fabrication of nanoscale porous heterostructures, where a single material is localized at the pore surface as a continuous layer. This SDA binding motif (design rule 1) enables sequence-controlled heterostructures, where the composition profile and interfaces correspond to the synthetic addition order. This approach is generalized with 5 material sequences using an SDA with only persistent SDA-NP interactions (“P-NP 1 –NP 2 ”; NP i = TiO 2 , Nb 2 O 5 , ZrO 2 ). Expanding these polymer SDA design guidelines, it is shown that the combination of both persistent and dynamic (noncovalent) SDA-NP interactions (“PD-NP 1 –NP 2 ”) improves the production of uniform interconnected porosity (design rule 2). The resulting competitive binding between two segments of the SDA (P- vs D-) requires additional time for the first NP type (NP 1 ) to reach and covalently attach to the SDA (design rule 3). The combination of these three design rules enables the direct self-assembly of heterostructures that localize a single material at the pore surface while preserving continuous porosity.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Emerging protein sequencing technologies: proteomics without mass spectrometry?

Liquid chromatography-tandem mass spectrometry (LC-MS/MS) has been a leading method for proteomics for 30 years. Advantages provided by LC-MS/MS are offset by significant disadvantages, including cost. Recently, several non-mass spectrometric methods have emerged, but little information is available about their capacity to analyze the complex mixtures routine for mass spectrometry. Areas Covered: We review recent non-mass-spectrometric methods for sequencing proteins and peptides, including those using nanopores, sequencing by degradation, reverse translation, and short-epitope mapping, with comments on bioinformatics challenges, fundamental limitations, and areas where new technologies will be more or less competitive with LC-MS/MS. In addition to conventional literature searches, instrument vendor websites, patents, webinars, and preprints were also consulted to give a more up-to-date picture. Expert Opinion: Many new technologies are promising. However, demonstrations that they outperform mass spectrometry in terms of peptides and proteins identified have not yet been published, and astute observers note important disadvantages, especially relating to the dynamic range of single-molecule measurements of complex mixtures. Still, even if the performance of emerging methods proves inferior to LC-MS/MS, their low cost could create a different kind of revolution: a dramatic increase in the number of biology laboratories engaging in new forms of proteomics research.

59 BASIC BIOLOGICAL SCIENCES↗

Dual transposon sequencing profiles the genetic interaction landscape in bacteria

Gene redundancy complicates systematic characterization of gene function as single-gene deletions may not produce discernible phenotypes. We report dual transposon sequencing (dual Tn-seq), a platform for assaying the fitness of a comprehensive double mutant pool in parallel. Dual Tn-seq couples random barcode transposon site sequencing with the Cre-lox system, enabling deep sampling of 73% of the 1.3 million possible double gene deletions in Streptococcus pneumoniae. The genetic interactions identified span a wide range of biochemical processes, revealing new factors in presumably well-studied pathways, exemplified by a cytidine triphosphate synthase PyrJ. Moreover, this approach should permit further investigation of growth condition–specific genetic interactions. Because dual Tn-seq does not require the construction of a large array of single mutants, it should be readily adaptable to various microorganisms.

CTP synthesis↗

An open control sequence specification to scale building demand flexibility via analytics software

For over two decades, researchers and practitioners have showcased the ability of large commercial buildings to provide grid services by shedding or shifting load. Various utility demand response (DR) and virtual power plant (VPP) programs throughout the United States are presently utilizing these demand-side resources. However, growth of these programs have been limited, in part due to the high cost necessary to integrate the DR control strategies into the building automation system (BAS). Implementing these strategies involves adjusting control sequences, necessitating dozens of hours of customized programming per building, limiting their adoption to large organizations and progressive owners. Recent efforts by researchers and industry have demonstrated the capability of energy management and information systems (EMIS), originally designed for fault detection and diagnostics, to interface with existing BAS and perform supervisory control to optimize building operations. While these approaches are quickly being adopted by industry, demand flexibility (DF) control strategies remain limited in product offerings. One of the challenges is the lack of documented best-practice DF sequences, despite the rich literature on field implementations. This paper develops a new open-specification for a zone-based temperature adjustment shed strategy for commercial building HVAC systems, describing the specification’s implementation in two EMIS tools in both experimental and field settings. Both implementations successfully reduced electric load by at least 40% on average during the called event, while maintaining temperature limits. This study’s detailed process from specification to deployment shows the potential for scalability as well as highlights challenges related to integration with heterogeneous BAS products.

Granderson, Jessica↗

MjCyc: Rediscovering the pathway-genome landscape of the first sequenced archaeon, Methanocaldococcus (Methanococcus) jannaschii

The genome of Methanocaldococcus (Methanococcus) jannaschii DSM 2661 was the first Archaeal genome to be sequenced in 1996. Subsequent sequence-based annotation cycles led to its first metabolic reconstruction in 2005. Leveraging new experimental results and function assignments, we have now re-annotated M. jannaschii, creating an updated resource with novel information and testable predictions in a pathway-genome database available at BioCyc.org. This reannotation effort has resulted in 652 function assignments with enzyme roles, accounting for a third of the total protein-coding entries for this genome. The updated resource includes 883 reactions, 540 enzymes, and 142 individual pathways. Despite notable progress in computational genomics, more than a third of the genome remains functionally uncharacterized. The publicly available MjCyc pathway-genome database holds great potential for the wider community to conduct research on the biology of methanogenic Archaea.

59 BASIC BIOLOGICAL SCIENCES↗

Hierarchical Chiral Self-Assembly of Nanocylinders Composed of Sequence-Defined Mesogenic Dimers

Chiral ensembles can arise through supramolecular curvature that resolves geometric frustrations in the packing of bent, achiral molecular or colloidal building blocks. Here, we leverage orthogonal protection−deprotection click chemistry to create sequence-defined mesogenic heterodimers exhibiting emergent chirality. We compare the hierarchical self-assembly of the synthesized asymmetric, achiral heterodimers, which differ only in the position of a methyl substituent. Both dimers form chiral spherulites composed of nanocylinders. However, the detailed arrangement of nanocylinders depends on the position of the methyl substituent and the crystallization conditions. Despite the chemical similarity, in one dimer, two crystalline forms are optically active. They form conglomerates of dextrorotatory and levorotatory spherulites. The other dimer forms more highly anisotropic spherulites that mask circular birefringence arising from the misorientation of nanocylinders, while mapping of nanocylinder directors reveals a sense at the spherulite surface. We propose that differences in nanocylinder arrangements may arise from changes in nanocylinder curvature and dimensions dictated by the methyl substituent position, inducing chirality. These results demonstrate multiscale hierarchical assembly relevant to dense systems of tubular structures and highlight the role of sequence and molecular design in directing the bottom-up hierarchical self-assembly and chirality of mesogenic systems.

Alkyls↗

Design of diverse, functional mitochondrial targeting sequences across eukaryotic organisms using variational autoencoder

Mitochondria play a key role in energy production and metabolism, making them a promising target for metabolic engineering and disease treatment. However, despite the known influence of passenger proteins on localization efficiency, only a few protein-localization tags have been characterized for mitochondrial targeting. To address this limitation, we leverage a Variational Autoencoder to design novel mitochondrial targeting sequences. In silico analysis reveals that a high fraction of the generated peptides (90.14%) are functional and possess features important for mitochondrial targeting. We characterize artificial peptides in four eukaryotic organisms and, as a proof-of-concept, demonstrate their utility in increasing 3-hydroxypropionic acid titers through pathway compartmentalization and improving 5-aminolevulinate synthase delivery by 1.62-fold and 4.76-fold, respectively. Moreover, we employ latent space interpolation to shed light on the evolutionary origins of dual-targeting sequences. Overall, our work demonstrates the potential of generative artificial intelligence for both fundamental research and practical applications in mitochondrial biology.

59 BASIC BIOLOGICAL SCIENCES↗

Ni-catalysed dicarbofunctionalization for the synthesis of sequence-encoded cyclooctene monomers

The properties of polymeric materials can be modulated by factors such as sequence control or functional group modifications. However, the synthesis of new macromolecular scaffolds is limited by the accessibility of structurally diverse monomers. This work describes a one-step, nickel-catalysed synthesis of 5,6-diaryl cyclooctene monomers from the feedstock chemical 1,5-cyclooctadiene. The reaction proceeds in a modular, regio- and diastereoselective fashion, granting access to both homo- and hetero-diaryl cyclooctene monomers that smoothly undergo ring-opening metathesis polymerization (ROMP). The resulting 1,2-diaryl-substituted polymers possess sequences with head-to-head styrene dyads that have not been previously explored, giving rise to unique and tunable properties. Density functional theory calculations highlight mechanistic aspects of the nickel-catalysed diarylation reaction and the ruthenium-catalysed ROMP process, revealing a previously unappreciated role of the boronic ester in promoting migratory insertion, which was leveraged to provide enantioinduction.

Catalytic mechanisms↗

ALS mutations disrupt self-association between the ubiquilin STI1 hydrophobic groove and internal placeholder sequences

Ubiquilins are molecular chaperones that play multifaceted roles in proteostasis, with point mutations in UBQLN2 leading to altered phase-separation properties and amyotrophic lateral sclerosis (ALS). Our mechanistic understanding of this essential process has been hindered by a lack of structural information on the STI1 domain, which is essential for ubiquilin chaperone activity and phase separation. Here, we present the first crystal structure of a ubiquilin-family STI1 domain bound to a transmembrane domain (TMD), and show that ALS mutations disrupt the STI1-TMD interaction. We further demonstrate that ubiquilins contain multiple conserved internal sequences that bind to the STI1 domain, including the PXX-repeat region that is a hotspot for ALS mutations. We propose that these placeholder sequences prevent solvent exposure of the STI1 hydrophobic groove and contribute to the multivalency that drives ubiquilin phase-separation. Together, this work provides a new paradigm for understanding how STI1 domains modulate ubiquilin chaperone activity and phase separation, and offers insights into the molecular basis of ALS pathogenesis.

Onwunma, Joan [Univ. of Toledo, OH (United States)↗

Discovering methylated DNA motifs in bacterial nanopore sequencing data with MIJAMP

Abstract Bacterial DNA methylation is involved in diverse cellular functions, including modulation of gene expression, DNA repair, and restriction–modification systems for defense against viruses and other foreign DNA. Restriction systems hinder efforts to engineer organisms to produce fuels and chemicals from waste and renewable feedstocks by degrading DNA during transformation. Methylome analysis allows identification of motifs within a bacterial chromosome that may be targeted by native restriction enzymes. Further expression of the corresponding methyltransferases in Escherichia coli allows plasmid DNA to be protected from restriction in the target organism, thereby drastically enhancing transformation efficiency. Nanopore sequencing can detect methylated bases, but software is needed to transform modified base coordinates into methylated motifs. Here, we develop MIJAMP (MIJAMP Is Just A MethylBED Parser), a software package that was developed to discover methylated motifs from the output of ONT’s Modkit or other data in the methylBED format. MIJAMP employs a human-driven refinement strategy that empirically validates all motifs against genome-wide methylation data, thus eliminating incorrect motifs. MIJAMP also reports methylation data on specific, user-defined motifs. Using MIJAMP, we determined the methylated motifs both in a control strain (wild-type E. coli) and in Synecococcus sp. strain PCC7002, laying the foundation for improved transformation in this organism. MIJAMP is available at https://code.ornl.gov/alexander-public/mijamp/. One Sentence Summary: Here we describe software written to discover DNA methylation motifs from nanopore sequencing data.

59 BASIC BIOLOGICAL SCIENCES↗

Integration of ultra-low coverage whole-genome sequences for reconstructing the evolutionary history of Galapagos giant tortoises

Genomic data from contemporary and historical samples often need to be coupled for evolutionary reconstructions of multitaxon complexes. However, the genetic data recovered from historical samples may result only in ultra-low coverage whole-genome sequences (ulcWGS; <0.15× depth), leading to inaccurate evolutionary inferences given a preponderance of missing data. Using the Galapagos giant tortoise radiation as a study system (Chelonoidis spp., composed of 13 extant and four extinct lineages), we assembled a novel methodological pipeline that removes potential noise introduced by the missing data and enhances the evolutionary signal from ulcWGS samples. We leveraged existing tools for phylogenomic placement (EPA-ng), population genomic structure (smartsnp) and admixture (Admixfrog, NGSadmix) to demonstrate that the evolutionary history of samples can be uncovered with sequencing depths as low as 0.008–0.139×. Importantly, these approaches do not use genotype imputation of the ulcWGS samples, which would require extensive reference datasets. Our application to two cases of extinct lineages of Galapagos giant tortoises, with and without references from the same lineage, demonstrates the general value of the approach. We confirm where the extinct lineages from San Cristóbal and Santa Fe islands fit into the Galapagos giant tortoise radiation, and that these lineages were evolutionarily distinct entities.

ancient DNA↗

Performance Evaluation of a Novel Sequence-Based Directional Detection Strategy for Protection of Active Distribution Networks

Directional elements are relied on to achieve selectivity in fault detection in power systems. Although such elements have been deployed successfully for many years, there is an increased need for novel methods to deal with the unique challenges of directional protection in modern distribution networks. This article analyzes the impact of inverter-based resources (IBRs) on existing directional protection methods in distribution systems. It identifies parts of such elements that pose a risk of misoperation when IBRs are used in distribution networks. The authors have developed a new directional detection method for unbalanced faults in such networks using superimposed symmetrical sequence quantities. The phase angle of the superimposed negative sequence admittance is used to determine fault direction. The paper also presents a real-time co-simulation platform between a simulated distribution system and physical protection relay, using OPAL-RT. An SEL-411L relay is used to program the detection algorithm. This hardware-in-the-loop (HIL) setup is used to verify the performance of the method and the results are compared with existing directional methods

24 POWER TRANSMISSION AND DISTRIBUTION↗

A Digital Three Level Space Vector Modulator for High Frequency Vector Sequence Generation

This letter proposes a digital high-speed three-level space vector pulse width modulator (3L-SVPWM). A conventional 3L-SVPWM is typically computation-based, involving a sequential execution of sub-tasks on a digital signal processor (DSP) based controller. The resulting high computation time of 5.4 μs limits the implementation of additional control blocks for switching frequencies greater than 100 kHz. This is overcome by transforming sub-tasks into digital blocks with 1-0 decisions and simpler arithmetic operations. The sub-task blocks are executed concurrently on a programmable logic device (PLD). Hence, a fast 3L-SVPWM execution in 140 ns is achieved. The proposed digital 3L-SVPWM enables high switching frequency operation of wide bandgap (WBG) device-based 3 L inverters to generate high fundamental frequency waveforms. A finite state machine is an integral part of the proposed implementation with the ability to generate any vector sequence, maximizing the usage of redundant vector states in 3L-SVPWM. Here, the proposed digital 3L-SVPWM operation is demonstrated with a GaN-based 3 L active neutral point clamped (3L-ANPC) inverter. Experimental results are presented at 250 kHz switching frequency to generate vector sequences for center-aligned SVPWM (CA-SVPWM) and common mode voltage reduced SVPWM (CMVR-SVPWM). The results also showcase a high fundamental frequency generation capability of 10 kHz.

active neutral point clamped inverter↗

Complete genome sequence of Luteolibacter sp. strain Populi, a member of phylum Verrucomicrobiota isolated from the Populus trichocarpa rhizosphere

Luteolibacter sp. strain Populi is a bacterium from the phylum Verrucomicrobiota, isolated from the rhizosphere of a black cottonwood tree, Populus trichocarpa, from the Cascade mountains in Washington. Its 6.6-Mb chromosome was completely sequenced using Oxford Nanopore long-read sequencing and is predicted to encode 5,301 proteins and 60 RNAs.

59 BASIC BIOLOGICAL SCIENCES↗