Search NASASearch

SEARCH · Search NASA

Results for “genome sequencing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Draft Genome Sequences of Biosafety Level 2 Opportunistic Pathogens Isolated From the Environmental Surfaces of the International Space Station

The draft genome sequences of 20 biosafety level 2 (BSL-2) opportunistic pathogens isolated from the environmental surfaces of the International Space Station (ISS) were presented. These genomic sequences will help in understanding the influence of microgravity on the pathogenicity and virulence of these strains when compared with Earth strains.

Aleksandra Checinska Sielaff

Analysis of the myosins encoded in the recently completed Arabidopsis thaliana genome sequence

BACKGROUND: Three types of molecular motors play an important role in the organization, dynamics and transport processes associated with the cytoskeleton. The myosin family of molecular motors move cargo on actin filaments, whereas kinesin and dynein motors move cargo along microtubules. These motors have been highly characterized in non-plant systems and information is becoming available about plant motors. The actin cytoskeleton in plants has been shown to be involved in processes such as transportation, signaling, cell division, cytoplasmic streaming and morphogenesis. The role of myosin in these processes has been established in a few cases but many questions remain to be answered about the number, types and roles of myosins in plants. RESULTS: Using the motor domain of an Arabidopsis myosin we identified 17 myosin sequences in the Arabidopsis genome. Phylogenetic analysis of the Arabidopsis myosins with non-plant and plant myosins revealed that all the Arabidopsis myosins and other plant myosins fall into two groups - class VIII and class XI. These groups contain exclusively plant or algal myosins with no animal or fungal myosins. Exon/intron data suggest that the myosins are highly conserved and that some may be a result of gene duplication. CONCLUSIONS: Plant myosins are unlike myosins from any other organisms except algae. As a percentage of the total gene number, the number of myosins is small overall in Arabidopsis compared with the other sequenced eukaryotic genomes. There are, however, a large number of class XI myosins. The function of each myosin has yet to be determined.

NASA Discipline Plant Biology

Congruence of Clusters Defined By Whole Genome Sequencing and MALDI-TOF for Bacteria Isolated From Cleanrooms

Introduction: Oligotrophic conditions can render cleanrooms inhospitable to microbes. Despite these constraints, fungi and bacteria are frequently isolated from surfaces in astromaterials cleanrooms at the Johnson Space Center. Bacillus species are of particular concern because endospores belonging to this genus are resilient and can affect astromaterials. Current monitoring programs rely on 16S rRNA sequencing and the VITEK2 Compact system. These methods have limited power to resolve Bacillus species. Matrix-assisted laser desorption - time of flight mass spectrometry (MALDI-TOF MS), provides a rapid, low cost, method of identifying bacterial isolates and has a higher resolution than 16S rRNA sequencing, particularly for Bacillus species; however, few studies have compared this method to the industry gold standard, whole genome sequencing (WGS). Methods: Based on 16S rRNA classification, we selected 14 isolates for analysis with MALDI-TOF and WGS. Mass spectra were generated with MALDI-TOF MS and processed with custom scripts to identify clusters of closely related isolates and calculate a matrix of pairwise cosine similarity scores. Hybrid Illumina and Nanopore sequencing were used to generate draft genomes. Pairwise similarity scores were calculated from these genomes based on the average amino acid identity (AAI) predicted from single copy core genes. Congruence of clustering between these methods, was assessed by calculating adjusted Rand and Wallace coefficients. Results: Clusters of species generated from MALDI-TOF MS showed good agreement of phylotypes generated with WGS. Pairs of strains that were > 94% similar to each other in terms of predicted amino acid sequences consistently showed cosine similarities of mass spectra > 0.65 and, of the 9 clusters identified with WGS, 8 were identical with MALDI-TOF. This corresponds to an adjusted Rand index of 0.95 and a 95% confidence interval of 0.80 – 1.00 for adjusted Wallace coefficients. The only discordance was for a pair of isolates that were classified as Paenibacillus species. This pair showed relatively high similarity (0.84) in terms of MALDI-TOF MS but only 85% similarity in terms of AAI. Conclusion: This study shows that MALDI-TOF and WGS exhibit a similar ability to delineate Bacillus species isolated from cleanrooms and taxonomic units described by these two methods are consistent with one another. Since MALDI-TOF MS is low in cost and high in throughput, this approach appears to be an ideal option for routine microbial monitoring and identifying Bacillus species.

Farnaz Mazhari

Draft Genome Sequences from a Novel Clade of Bacillus cereus Sensu Lato Strains, Isolated from the International Space Station

The draft genome sequences of six Bacillus strains, isolated from the International Space Station and belonging to the Bacillus anthracis-B. cereus-B. thuringiensis group, are presented here. These strains were isolated from the Japanese Experiment Module (one strain), U.S. Harmony Node 2 (three strains), and Russian Segment Zvezda Module (two strains).

Kasthuri Venkateswaran

Application of Whole Genome Sequencing and MALDI-TOF to the Identifcation of Bacillus Species Isolated from Cleanrooms at NASA Johnson Space Center

Astromaterial cleanrooms at NASA Johnson Space Center are built environments that hold samples, such as lunar rocks, from different space exploration missions. Bacillus sp. are frequently detected in routine microbial monitoring of these facilities. Since this, and related genera, can form endospores that can withstand harsh conditions, they could contaminate astromaterials. This could confound searches for extraterrestrial life. Whole genome sequencing (WGS) is widely used for identifying bacterial strains and tracking their source; however, WGS is expensive and time consuming. Matrix-assisted laser desorption ionization– time of flight mass spectrometry (MALDI-TOF) shows promise as a low-cost, rapid method of identifying strains of bacteria, but few studies have compared this proteomics method to WGS. To evaluate a high throughput method of tracking the source of contamination of this built environment, WGS and MALDI-TOF was conducted on 18 bacterial strains isolated from surfaces in astromaterials cleanrooms. WGS identified 14 Bacillus, 2 Paenibacillus, 1 Solibacillus and 1 Alcaligenes strains. These isolates showed similarity to strains commonly observed in spacecraft assembly cleanrooms at other facilities. Cluster analysis of mass spectra generated by MALDI-TOF grouped strains together that were greater than 94% similar to each other in terms of amino acid sequences of single copy core genes, as assessed by WGS. This suggests that MALDI-TOF and WGS results are consistent with each other and MALDI-TOF can rapidly identify strains of Bacillus sp. isolated from cleanroom environments with a resolution comparable to WGS. Based on phylogenomic analysis, these results also suggest the presence of a cosmopolitan class of Bacillus sp. that are more likely to be found in cleanrooms and similar built environments than in natural systems.

Farnaz Mazhari

Genome Sequences of Bacteria Isolated from the International Space Station Water Systems

We report draft genomes of five bacteria recovered from the United States and Russian water systems onboard the International Space Station: bacteria of the genera Ralstonia, Burkholderia, Cupriavidus, Methylobacterium, and Pseudomonas. These sequences will help further the understanding of water reclamation and environmental control and life support systems in space.

Christian L. Castro

Resolution of Maldi-Tof Compared to Whole Genome Sequencing for Identification of Bacillus Species Isolated From Cleanrooms at Nasa Johnson Space Center

The Astromaterials Acquisition and Curation Office at NASA Johnson Space Center maintains cleanrooms to archive extraterrestrial materials returned from space exploration missions. Compared to typical built environments, oligotrophic conditions make these facilities inhospitable to microbes. Despite these controls, bacteria and fungi are regularly cultured from these cleanrooms. In particular, Bacillus sp. are frequently isolated during routine microbial monitoring. Endospores associated with this genus can survive extreme environments, such as cleanrooms. This microbial contamination may affect the integrity of astromaterials.

Microbiology

Molecular motors and their functions in plants

Molecular motors that hydrolyze ATP and use the derived energy to generate force are involved in a variety of diverse cellular functions. Genetic, biochemical, and cellular localization data have implicated motors in a variety of functions such as vesicle and organelle transport, cytoskeleton dynamics, morphogenesis, polarized growth, cell movements, spindle formation, chromosome movement, nuclear fusion, and signal transduction. In non-plant systems three families of molecular motors (kinesins, dyneins, and myosins) have been well characterized. These motors use microtubules (in the case of kinesines and dyneins) or actin filaments (in the case of myosins) as tracks to transport cargo materials intracellularly. During the last decade tremendous progress has been made in understanding the structure and function of various motors in animals. These studies are yielding interesting insights into the functions of molecular motors and the origin of different families of motors. Furthermore, the paradigm that motors bind cargo and move along cytoskeletal tracks does not explain the functions of some of the motors. Relatively little is known about the molecular motors and their roles in plants. In recent years, by using biochemical, cell biological, molecular, and genetic approaches a few molecular motors have been isolated and characterized from plants. These studies indicate that some of the motors in plants have novel features and regulatory mechanisms. The role of molecular motors in plant cell division, cell expansion, cytoplasmic streaming, cell-to-cell communication, membrane trafficking, and morphogenesis is beginning to be understood. Analyses of the Arabidopsis genome sequence database (51% of genome) with conserved motor domains of kinesin and myosin families indicates the presence of a large number (about 40) of molecular motors and the functions of many of these motors remain to be discovered. It is likely that many more motors with novel regulatory mechanisms that perform plant-specific functions are yet to be discovered. Although the identification of motors in plants, especially in Arabidopsis, is progressing at a rapid pace because of the ongoing plant genome sequencing projects, only a few plant motors have been characterized in any detail. Elucidation of function and regulation of this multitude of motors in a given species is going to be a challenging and exciting area of research in plant cell biology. Structural features of some plant motors suggest calcium, through calmodulin, is likely to play a key role in regulating the function of both microtubule- and actin-based motors in plants.

Non-NASA Center

Permanent Draft Genome of Strain ESFC-1: Ecological Genomics of a Newly Discovered Lineage of Filamentous Diazotrophic Cyanobacteria

The nonheterocystous filamentous cyanobacterium, strain ESFC-1, is a recently described member of the order Oscillatoriales within the Cyanobacteria. ESFC-1 has been shown to be a major diazotroph in the intertidal microbial mat system at Elkhorn Slough, CA, USA. Based on phylogenetic analyses of the 16S RNA gene, ESFC-1 appears to belong to a unique, genus-level divergence; the draft genome sequence of this strain has now been determined. Here we report features of this genome as they relate to the ecological functions and capabilities of strain ESFC-1. The 5,632,035 bp genome sequence encodes 4914 protein-coding genes and 92 RNA genes. One striking feature of this cyanobacterium is the apparent lack of either uptake or bi-directional hydrogenases typically expected within a diazotroph. Additionally, a large genomic island is found that contains numerous low GC-content genes and genes related to extracellular polysaccharide production and cell wall synthesis and maintenance.

R Craig Everroad

Discovery of Chlorophyll d : Isolation and Characterization of a Far-Red Cyanobacterium from the Original Site of Manning and Strain (1943) at Moss Beach, California

We have isolated a chlorophyll-d-containing cyanobacterium from the intertidal field site at Moss Beach, on the coast of Central California, USA, where Manning and Strain (1943) originally discovered this far-red chlorophyll. Here, we present the cyanobacterium’s environmental description, culturing procedure, pigment composition, ultrastructure, and full genome sequence. Among cultures of far-red cyanobacteria obtained from red algae from the same site, this strain was an epiphyte on a brown macroalgae. Its Q y in vivo absorbance peak is centered at 704–705 nm, the shortest wavelength observed thus far among the various known Acaryochloris strains. Its Chl a /Chl d ratio was 0.01, with Chl d accounting for 99% of the total Chl d and Chl a mass. TEM imagery indicates the absence of phycobilisomes, corroborated by both pigment spectra and genome analysis. The Moss Beach strain codes for only a single set of genes for producing allophycocyanin. Genomic sequencing yielded a 7.25 Mbp circular chromosome and 10 circular plasmids ranging from 16 kbp to 394 kbp. We have determined that this strain shares high similarity with strain S15, an epiphyte of red algae, while its distinct gene complement and ecological niche suggest that this strain could be the closest known relative to the original Chl d source of Manning and Strain (1943). The Moss Beach strain is designated Acaryochloris sp. (marina) strain Moss Beach.

chlorophyll d

Gene Fusion: A Genome Wide Survey

As a well known fact, organisms form larger and complex multimodular (composite or chimeric) and mostly multi-functional proteins through gene fusion of two or more individual genes which have independent evolution histories and functions. We call each of these components a module. The existence of multimodular proteins may improves the efficiency in gene regulation and in cellular functions, and thus may give the host organism advantages in adaptation to environments. Analysis of all gene fusions in present-day organisms should allow us to examine the patterns of gene fusion in context with cellular functions, to trace back the evolution processes from the ancient smaller and uni-functional proteins to the present-day larger and complex multi-functional proteins, and to estimate the minimal number of ancestor proteins that existed in the last common ancestor for all life on earth. Although many multimodular proteins have been experimentally known, identification of gene fusion events systematically at genome scale had not been possible until recently when large number of completed genome sequences have been becoming available. In addition, technical difficulties for such analysis also exist due to the complexity of this biological and evolutionary process. We report from this study a new strategy to computationally identify multimodular proteins using completed genome sequences and the results surveyed from 22 organisms with the data from over 40 organisms to be presented during the meeting. Additional information is contained in the original extended abstract.

Liang, Ping

Spaceflight Autonomous Multigenerational Microbial Sequencer (SAMMS) in Support of Plant-Growth Systems

As the National Aeronautics and Space Association (NASA) begins to pursue long-duration space flights, they will need to be able to provide astronauts with a nutritious and reliable food source. To meet the administration’s goal of traveling to the Moon and Mars, astronauts will need to begin to grow their own food in space. To protect their food source, extensive monitoring will occur to test for the effects of a space flight environment (e.g., radiation) as well as for early pathogen and disease detection. Genomic sequencing allows for both concerns to be tested on a regular basis. However, NASA’s current sequencer is unable to process plant tissues. Therefore, a novel method for plant DNA extraction using microneedle (MN) patches that will be able to feed into NASA’s existing system, but also require minimal human input is proposed. To support this, the design was broken down into four components (1) MN patch fabrication (2) MN patch extraction, (3) automated sampling motion control, and (4) a processing module. The MN patch is fabricated using a custom mold with conically shaped needles. The mold is filled with Polyvinyl alcohol (PVA) solution and placed in a vacuum desiccator. The mold is left in the vacuum overnight until the patch is dry and ready for use. The protocol was tested with varying pressures, drying times, volume amounts, and preparation methods to determine if highquality needles can be produced. A MN is a method of DNA extraction where the patch is applied to a leaf, the needles penetrate the leaf, breaking the rigid plant cell wall to isolate the DNA. A protocol for this method of extraction was tested to ensure the patch could produce the needed yield and purity. The tests varied by the number of patches, number of applications, and plant type. To automate the MN extraction method, motion control will utilize two separate axis tables which move in the x and y directions. The y-axis table will have an end effector that fits a MN patch and will have the ability to apply the patch to the leaf sample. This end effector will also act as a lid for a downstream processing module. The other axis will position the leaf sample and processing container so that the patch can be applied accurately. The Joint Comprehensive Sequencing System (JCSS) module integrates all the components together. The output of this module feeds into the NASA Charged Information-Storage Polymer Preparation System (CHIPPS) for genomic sequencing. The extraction module operates using a series of syringes and tubing to pump the varying reagents needed for the extraction protocol. The results of the study proved that MN patches are a viable method of DNA extraction. While fabrication of high-quality needles was unsuccessful, the protocol was able to be further developed using centrifugation. The integrated design between the motion control and the JCSS enabled the potential for automation with a complete conceptual design and prototype. Future research and development for this study would include (1) further testing for fabrication (2) expanding the range of plant species compatible with the MN patch, and (3) building a working prototype for the integrated system.

Peter Ling

Life in the Fast Lane for Protein Crystallization and X-Ray Crystallography

The common goal for structural genomic centers and consortiums is to decipher as quickly as possible the three-dimensional structures for a multitude of recombinant proteins derived from known genomic sequences. Since X-ray crystallography is the foremost method to acquire atomic resolution for macromolecules, the limiting step is obtaining protein crystals that can be useful of structure determination. High-throughput methods have been developed in recent years to clone, express, purify, crystallize and determine the three-dimensional structure of a protein gene product rapidly using automated devices, commercialized kits and consolidated protocols. However, the average number of protein structures obtained for most structural genomic groups has been very low compared to the total number of proteins purified. As more entire genomic sequences are obtained for different organisms from the three kingdoms of life, only the proteins that can be crystallized and whose structures can be obtained easily are studied. Consequently, an astonishing number of genomic proteins remain unexamined. In the era of high-throughput processes, traditional methods in molecular biology, protein chemistry and crystallization are eclipsed by automation and pipeline practices. The necessity for high rate production of protein crystals and structures has prevented the usage of more intellectual strategies and creative approaches in experimental executions. Fundamental principles and personal experiences in protein chemistry and crystallization are minimally exploited only to obtain "low-hanging fruit" protein structures. We review the practical aspects of today s high-throughput manipulations and discuss the challenges in fast pace protein crystallization and tools for crystallography. Structural genomic pipelines can be improved with information gained from low-throughput tactics that may help us reach the higher-bearing fruits. Examples of recent developments in this area are reported from the efforts of the Southeast Collaboratory for Structural Genomics (SECSG).

Pusey, Marc L.

Life in the fast lane for protein crystallization and X-ray crystallography

The common goal for structural genomic centers and consortiums is to decipher as quickly as possible the three-dimensional structures for a multitude of recombinant proteins derived from known genomic sequences. Since X-ray crystallography is the foremost method to acquire atomic resolution for macromolecules, the limiting step is obtaining protein crystals that can be useful of structure determination. High-throughput methods have been developed in recent years to clone, express, purify, crystallize and determine the three-dimensional structure of a protein gene product rapidly using automated devices, commercialized kits and consolidated protocols. However, the average number of protein structures obtained for most structural genomic groups has been very low compared to the total number of proteins purified. As more entire genomic sequences are obtained for different organisms from the three kingdoms of life, only the proteins that can be crystallized and whose structures can be obtained easily are studied. Consequently, an astonishing number of genomic proteins remain unexamined. In the era of high-throughput processes, traditional methods in molecular biology, protein chemistry and crystallization are eclipsed by automation and pipeline practices. The necessity for high-rate production of protein crystals and structures has prevented the usage of more intellectual strategies and creative approaches in experimental executions. Fundamental principles and personal experiences in protein chemistry and crystallization are minimally exploited only to obtain "low-hanging fruit" protein structures. We review the practical aspects of today's high-throughput manipulations and discuss the challenges in fast pace protein crystallization and tools for crystallography. Structural genomic pipelines can be improved with information gained from low-throughput tactics that may help us reach the higher-bearing fruits. Examples of recent developments in this area are reported from the efforts of the Southeast Collaboratory for Structural Genomics (SECSG).

Review

Cleanroom Microbes Survive Drying, Vacuum, and Proton Irradiation

Introduction : The goal of planetary protection at NASA is to mitigate the risk of contaminating sensitive target bodies with biological life. While many cleaning procedures have been put in place to reduce bioburden on spacecraft, microbes are experts at evolving to survive harsh conditions. Specifically, the dry, low-nutrient environment of a cleanroom (commonly used for assembly of spacecraft) can represent an environment where extremophiles can survive. Methods : Scientists at NASA MSFC wished to gather a snapshot of the microbial population within a variety of cleanrooms on site. A study was undertaken to collect air, surface, and floor samples from clean-rooms and isolate unique morphologies. From this study, 95 isolates were collected and saved in a microbial library. About 86% of these were identified at least to a genus level. Following identification, 24 microbes were selected, based on a literature review, as potential extremophiles. These were grown in liquid cultures, diluted to a set optical density, washed with water, and then applied to a sterilized Kapton coupon. Droplets were allowed to dry overnight in a biosafety cabinet. Coupons were then installed in a pelletron and pumped down to high vacuum (~1E-6 Torr). Samples were then subjected 100 keV protons at a fluence of 2x10 15 p+/cm 2 up to 4x10 15 p+/cm 2 . Following exposure, samples were returned to the microbiology lab where they were pro-cessed by submerging in water, vortexing, and then plating either droplets or spread plates. Recovery data collected was qualitative with a ranking or +, minor, or – for growth. Some selected radiotolerant strains were sequenced using the Illumina sequencing platform. The resulting genomes were annotated with the Rapid Annotations using Subsystems Technology (RAST) server and analyzed for conserved and unique stress response relevant genomic signatures to identify clues related to specific tolerances. Results and Discussion : After five rounds of proton radiation, we narrowed our isolates to five, non-spore forming bacteria that demonstrated survival: Arthrobacter koreensis, Paenarthrobacter nitroguajacolicus, Mycetocola manganoxydans , and an Erwinia sp. Furthermore, we exposed these four microbes to 254 nm wavelength light at an intensity of 80 W/m 2 at a distance of ~18 cm for 10 minutes. Only A. koreensis demonstrated survival following UV exposure. Finally, we performed whole genome sequencing on the four strains to look for genetic markers of stress resistance. When we compared the genomes of the four strains, we found that genes coding for GGDEF and EAL domains with PAS/PAC sensors were only found in A. koreensis . These domains, modulated by PAS/PAC sensors, are hypothesized to facilitate survival under drying, desiccation, and proton irradiation. Drying and Desiccation : PAS domains sense hydration changes and modulate GGDEF and EAL domain activity to adjust c-di-GMP levels, enhancing resistance to desiccation. For instance, in Pseudomonas aeruginosa , the PAS domain of RbdA modulates activity under varying hydration conditions, affecting stress responses [1]. Proton Irradiation : Proton irradiation causes oxidative stress, leading to ROS generation. PAS domains detect this stress and modulate GGDEF and EAL domains to manage oxidative stress responses. In Shewanella , EAL domain proteins modulated by PAS sensors help bacteria adapt to extreme conditions [2]. These genes upregulate other stress response genes, protecting membrane function, protein stability, DNA repair, and antioxidant defenses. The modulation of c-di-GMP by PAS domains is crucial for bacterial adaptation to stress conditions, enabling dynamic physio-logical adjustments [3]. Understanding these mechanisms provides insights into bacterial stress responses and strategies for controlling bacterial growth [4]. Conclusions : These findings indicate that clean-rooms harbor extremophile microbes that may be able to survive conditions in deep space. Furthermore, while we identified certain stress-response genes that may be at least partly responsible for the phenotypes observed in this study, there are likely unidentified genes or characteristics about A. koreensis , and other bacteria, that may allow them to survive in harsh environments. Future studies will focus on identifying these unknown genes and characteristics, further elucidating the mechanisms of extremophile survival and potentially informing the development of new biotechnologies for space exploration and other extreme environments.

Chelsi Cassilly

Developing Open-Source Training Materials for AI/ML and Space Biological Sciences Using NASA Cloud-Based Data

Artificial Intelligence (AI) and Machine Learning (ML) has gained significant traction in the biological and biomedical research fields in the last two decades, in part thanks to an increasing culture of open data sharing and reuse. Due to its capability for identifying complex relationships and patterns, AI/ML methodology is particularly well suited to recognize and predict biological patterns from high-dimensional next-generation sequencing data (e.g. whole genome sequencing, transcriptomic sequencing), as well as from biological or medical imaging data (e.g. microscopy, computed tomography, ultrasound, magnetic resonance imaging, radiography). These methodologies hold particular promise for space biosciences research and automated space health monitoring systems. However, there are many key considerations for properly training, validating, and testing a machine learning model in biological research or clinical application. Even with the positive culture of Open Science and data sharing, inexperienced researchers working quickly without proper checks can produce models that perform poorly outside of the immediate training dataset. Lessons learned from biological AI/ML research indicate that Open Science principles such as data sharing and open-source code must go hand-in-hand with publicly available, high-quality training curricula in best practices, with modules centered on real-life scientific use cases and data so future AI/ML practitioners gain experience on real problems. Here we present the development of open-source training materials for AI/ML and space biosciences, as part of the NASA Transform to Open Science Training (TOPST) initiative. We develop 4 independent training programs, focused on the following topics: 1) Fundamentals of Machine Learning and Space Biosciences Domain, 2) Open Science, Artificial Intelligence, and Ethical Best Practices for Data Sharing and Analysis, 3) Using AI/ML Classification to Identify Gene Networks Affected By Space Exposure in Mouse Liver, and 4) Using Neural Networks to Find DNA Damage Patterns in Immune Cells after Radiation. All programs leverage cloud-based NASA biological datasets. The curriculum we present will enable worldwide access to training in AI/ML and scientific analysis.

James Andrew Casaletto