Search NASA⌕ Search

SEARCH · Search NASA

Results for “genomics”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 289 records · Page 16

Discovering type I cis-AT polyketides through computational mass spectrometry and genome mining with Seq2PKS

Type 1 polyketides are a major class of natural products used as antiviral, antibiotic, antifungal, antiparasitic, immunosuppressive, and antitumor drugs. Analysis of public microbial genomes leads to the discovery of over sixty thousand type 1 polyketide gene clusters. However, the molecular products of only about a hundred of these clusters are characterized, leaving most metabolites unknown. Characterizing polyketides relies on bioactivity-guided purification, which is expensive and time-consuming. To address this, we present Seq2PKS, a machine learning algorithm that predicts chemical structures derived from Type 1 polyketide synthases. Seq2PKS predicts numerous putative structures for each gene cluster to enhance accuracy. The correct structure is identified using a variable mass spectral database search. Benchmarks show that Seq2PKS outperforms existing methods. Applying Seq2PKS to Actinobacteria datasets, we discover biosynthetic gene clusters for monazomycin, oasomycin A, and 2-aminobenzamide-actiphenol.

60 APPLIED LIFE SCIENCES↗

High allelic diversity in Arabidopsis NLRs is associated with distinct genomic features

Plants rely on Nucleotide-binding, Leucine-rich repeat Receptors (NLRs) for pathogen recognition. Highly variable NLRs (hvNLRs) show remarkable intraspecies diversity, while their low-variability paralogs (non-hvNLRs) are conserved between ecotypes. At a population level, hvNLRs provide new pathogen-recognition specificities, but the association between allelic diversity and genomic and epigenomic features has not been established. Our investigation of NLRs in Arabidopsis Col-0 has revealed that hvNLRs show higher expression, less gene body cytosine methylation, and closer proximity to transposable elements than non-hvNLRs. hvNLRs show elevated synonymous and nonsynonymous nucleotide diversity and are in chromatin states associated with an increased probability of mutation. Diversifying selection maintains variability at a subset of codons of hvNLRs, while purifying selection maintains conservation at non-hvNLRs. How these features are established and maintained, and whether they contribute to the observed diversity of hvNLRs is key to understanding the evolution of plant innate immune receptors.

59 BASIC BIOLOGICAL SCIENCES↗

Metagenome‐Assembled Genomes for Oligotrophic Nitrifiers From a Mountainous Gravelbed Floodplain

Riparian floodplains are important regions for biogeochemical cycling, including nitrogen. Here, we present MAGs from nitrifying microorganisms, including ammonia-oxidising archaea (AOA) and comammox bacteria from Slate River (SR) floodplain sediments (Crested Butte, CO, US). Additionally, we explore MAGs from potential nitrite-oxidising bacteria (NOB) from the Nitrospirales. AOA diversity in SR is lower than observed in other western US floodplain sediments and Nitrosotalea-like lineages such as the genus TA-20 are the dominant AOA. No ammonia-oxidising bacteria (AOB) MAGs were recovered. Microorganisms from the Palsa-1315 genus (clade B comammox) are the most abundant ammonia-oxidizers in SR floodplain sediments. Established NOB are conspicuously absent; however, we recovered MAGs from uncultured lineages of the NS-4 family (Nitrospirales) and Nitrospiraceae that we propose as putative NOB. Nitrite oxidation may be carried out by organisms sister to established Nitrospira NOB lineages based on the genomic content of uncultured Nitrospirales clades. Nitrifier MAGs recovered from SR floodplain sediments harbour genes for using alternative sources of ammonia, such as urea, cyanate, biuret, triuret and nitriles. In conclusion, the SR floodplain therefore appears to be a low ammonia flux environment that selects for oligotrophic nitrifiers.

60 APPLIED LIFE SCIENCES↗

Genome‐wide association identifies a BAHD acyltransferase activity that assembles an ester of glucuronosylglycerol and phenylacetic acid

SUMMARY Genome‐wide association studies (GWAS) are an effective approach to identify new specialized metabolites and the genes involved in their biosynthesis and regulation. In this study, GWAS of Arabidopsis thaliana soluble leaf and stem metabolites identified alleles of an uncharacterized BAHD‐family acyltransferase (AT5G57840) associated with natural variation in three structurally related metabolites. These metabolites were esters of glucuronosylglycerol, with one metabolite containing phenylacetic acid as the acyl component of the ester. Knockout and overexpression of AT5G57840 in Arabidopsis and heterologous overexpression in Nicotiana benthamiana and Escherichia coli demonstrated that it is capable of utilizing phenylacetyl‐CoA as an acyl donor and glucuronosylglycerol as an acyl acceptor. We, thus, named the protein Glucuronosylglycerol Ester Synthase (GGES). Additionally, phenylacetyl glucuronosylglycerol increased in Arabidopsis CYP79A2 mutants that overproduce phenylacetic acid and was lost in knockout mutants of UDP‐sulfoquinovosyl: diacylglycerol sulfoquinovosyl transferase, an enzyme required for glucuronosylglycerol biosynthesis and associated with glycerolipid metabolism under phosphate‐starvation stress. GGES is a member of a well‐supported clade of BAHD family acyltransferases that arose by duplication and neofunctionalized during the evolution of the Brassicales within a larger clade that includes HCT as well as enzymes that synthesize other plant‐specialized metabolites. Together, this work extends our understanding of the catalytic diversity of BAHD acyltransferases and uncovers a pathway that involves contributions from both phenylalanine and lipid metabolism.

09 BIOMASS FUELS↗

Complete genomes of Mucilaginibacter sabulilitoris SNA2 and Mucilaginibacter sp. cycad4: microbes with the potential for plant growth promotion

Mucilaginibacter species have been isolated from various environments, often in association with plants. Here, we report the complete genomes of Mucilaginibacter sabulilitoris SNA2 and Mucilaginibacter sp. cycad4. The former is the first available for that species, and based on 16S sequence analysis, the latter strain is likely a new species.

Mucilaginibacter↗

Metagenome-assembled genomes from Wind River Basin floodplain sediments Riverton, Wyoming site (August 2015)

Microorganisms play a key role in cycling nutrients and contaminants in the terrestrial environment depending on their genetic potential. Here we present metagenome-assembled genomes (MAGs) for the bacterial and archaeal community in floodplain sediment samples taken August 29, 2015 at a location (KB1) close to DOE Legacy Management well 855 at the Riverton, Wyoming floodplain site in the Wind River Basin (WRB). The groundwater at this site exhibits persistent U, Mo, and sulfate plumes and is one of the field sites in focus for the SLAC Groundwater Quality SFA program. Sediment samples from a deep soil pit were collected from 0 to 234 cm depth below surface at discrete depths every ~10-20 cm for microbial analyses. 13 metagenomes were sequenced through JGI and can be found under Gold sequencing project: Gs0131241. Metagenomes were assembled, binned, and refined using metawrap to generate MAGs (>50% complete and < 10% contamination based on checkM scores).This dataset includes a zip file of 2216 MAG fasta files and a csv file with quality, taxonomic classification (GTDB RS220), and metagenome accessions for MAGs. This dataset also includes a file-level metadata (flmd.csv) file that lists each file contained in the dataset with associated metadata and a data dictionary (dd.csv) file that contains column/row headers used throughout the files along with a definition, units, and data type

54 ENVIRONMENTAL SCIENCES↗

Metagenome-assembled genomes from Wind River Basin floodplain sediments Riverton, Wyoming site (June to October 2019)

Microorganisms play a key role in cycling nutrients and contaminants in the terrestrial environment depending on their genetic potential. Here we present metagenome-assembled genomes (MAGs) for the bacterial and archaeal community in floodplain sediment samples taken at three time points from June 12, 2019 to October 23,2019 at a location (PTT1) close to DOE Legacy Management well 855 at the Riverton, Wyoming floodplain site in the Wind River Basin (WRB). The groundwater at this site exhibits persistent U, Mo, and sulfate plumes and is one of the field sites in focus for the SLAC Groundwater Quality SFA program. Sediment samples were collected from 60 to 180 cm below surface every 30cm for microbial analyses through metagenomic sequencing. 15 metagenomes were sequenced through JGI and can be found under Gold sequencing project: Gs0131241. Metagenomes were assembled, binned, and refined using metawrap to generate MAGs (>50% complete and < 10% contamination based on checkM scores). This dataset includes a zip file of 780 MAG fasta files and a csv file with quality, taxonomic classification (GTDB RS220), and metagenome accessions for MAGs. This dataset also includes a file-level metadata (flmd.csv) file that lists each file contained in the dataset with associated metadata and a data dictionary (dd.csv) file that contains column/row headers used throughout the files along with a definition, units, and data type. A sample metadata file (samples.csv) that contains site information has also been included.

54 ENVIRONMENTAL SCIENCES↗

Metagenome-assembled genomes from Wind River Basin floodplain sediments Riverton, Wyoming site (May to September 2017)

Microorganisms play a key role in cycling nutrients and contaminants in the terrestrial environment depending on their genetic potential. Here we present metagenome-assembled genomes (MAGs) for the bacterial and archaeal community in floodplain sediment samples taken roughly every month in the period May 18 to September 13 in 2017 at a location (Pit2) close to DOE Legacy Management well 855 at the Riverton, Wyoming floodplain site in the Wind River Basin (WRB). The groundwater at this site exhibits persistent U, Mo, and sulfate plumes and is one of the field sites in focus for the SLAC Groundwater Quality SFA program. Cores were taken with a hand-auger and separated into 5-20 cm segments based on soil horizonation down to 150 cm depth below surface. Each segment was subsampled for microbial analyses. Corresponding 16S rRNA gene amplicon data is available at the NCBI Single Read Archive (SRA) Database BioProject ID PRJNA626616, and soil geochemistry data at doi:10.15485/1631972. 40 metagenomes were sequenced through JGI and can be found under Gold sequencing project: Gs0142591. Metagenomes were assembled, binned, and refined using metawrap to generate MAGs (>50% complete and < 10% contamination based on checkM scores). This dataset includes a zip file of 6993 MAG fasta files and a csv file with quality, taxonomic classification (GTDB RS220), and metagenome accessions for MAGs generated from the Wind River Basin (WRB). This dataset also includes a file-level metadata (flmd.csv) file that lists each file contained in the dataset with associated metadata and a data dictionary (dd.csv) file that contains column/row headers used throughout the files along with a definition, units, and data type.

54 ENVIRONMENTAL SCIENCES↗

Probing interspecies metabolic interactions within a synthetic binary microbiome using genome-scale modeling

Metabolic interactions within a microbial community play a key role in determining the structure, function, and composition of the community. However, due to the complexity and intractability of natural microbiomes, limited knowledge is available on interspecies interactions within a community. In this work, using a binary synthetic microbiome, a methanotroph-photoautotroph (M-P) coculture, as the model system, we examined different genome-scale metabolic modeling (GEM) approaches to gain a better understanding of the metabolic interactions within the coculture, how they contribute to the enhanced growth observed in the coculture, and how they evolve over time. Using batch growth data of the model M-P coculture, we compared three GEM approaches for microbial communities. Two of the methods are existing approaches: SteadyCom, a steady state GEM, and dynamic flux balance analysis (DFBA) Lab, a dynamic GEM. We also proposed an improved dynamic GEM approach, DynamiCom, for the M-P coculture. SteadyCom can predict the metabolic interactions within the coculture but not their dynamic evolutions; DFBA Lab can predict the dynamics of the coculture but cannot identify interspecies interactions. DynamiCom was able to identify the cross-fed metabolite within the coculture, as well as predict the evolution of the interspecies interactions over time. A new dynamic GEM approach, DynamiCom, was developed for a model M-P coculture. Constrained by the predictions from a validated kinetic model, DynamiCom consistently predicted the top metabolites being exchanged in the M-P coculture, as well as the establishment of the mutualistic N-exchange between the methanotroph and cyanobacteria. The interspecies interactions and their dynamic evolution predicted by DynamiCom are supported by ample evidence in the literature on methanotroph, cyanobacteria, and other cyanobacteria-heterotroph cocultures.

59 BASIC BIOLOGICAL SCIENCES↗

PyPop: a mature open-source software pipeline for population genomics

Python for Population Genomics (PyPop) is a software package that processes genotype and allele data and performs large-scale population genetic analyses on highly polymorphic multi-locus genotype data. In particular, PyPop tests data conformity to Hardy-Weinberg equilibrium expectations, performs Ewens-Watterson tests for selection, estimates haplotype frequencies, measures linkage disequilibrium, and tests significance. Standardized means of performing these tests is key for contemporary studies of evolutionary biology and population genetics, and these tests are central to genetic studies of disease association as well. Here, we present PyPop 1.0.0, a new major release of the package, which implements new features using the more robust infrastructure of GitHub, and is distributed via the industry-standard Python Package Index. New features include implementation of the asymmetric linkage disequilibrium measures and, of particular interest to the immunogenetics research communities, support for modern nomenclature, including colon-delimited allele names, and improvements to meta-analysis features for aggregating outputs for multiple populations.

59 BASIC BIOLOGICAL SCIENCES↗

Potential applications of microbial genomics in nuclear non-proliferation

As nuclear technology evolves in response to increased demand for diversification and decarbonization of the energy sector, new and innovative approaches are needed to effectively identify and deter the proliferation of nuclear arms, while ensuring safe development of global nuclear energy resources. Preventing the use of nuclear material and technology for unsanctioned development of nuclear weapons has been a long-standing challenge for the International Atomic Energy Agency and signatories of the Treaty on the Non-Proliferation of Nuclear Weapons. Environmental swipe sampling has proven to be an effective technique for characterizing clandestine proliferation activities within and around known locations of nuclear facilities and sites. However, limited tools and techniques exist for detecting nuclear proliferation in unknown locations beyond the boundaries of declared nuclear fuel cycle facilities, representing a critical gap in non-proliferation safeguards. Microbiomes, defined as “characteristic communities of microorganisms” found in specific habitats with distinct physical and chemical properties, can provide valuable information about the conditions and activities occurring in the surrounding environment. Microorganisms are known to inhabit radionuclide-contaminated sites, spent nuclear fuel storage pools, and cooling systems of water-cooled nuclear reactors, where they can cause radionuclide migration and corrosion of critical structures. Microbial transformation of radionuclides is a well-established process that has been documented in numerous field and laboratory studies. These studies helped to identify key bacterial taxa and microbially-mediated processes that directly and indirectly control the transformation, mobility, and fate of radionuclides in the environment. Expanding on this work, other studies have used microbial genomics integrated with machine learning models to successfully monitor and predict the occurrence of heavy metals, radionuclides, and other process wastes in the environment, indicating the potential role of nuclear activities in shaping microbial community structure and function. Results of this previous body of work suggest fundamental geochemical-microbial interactions occurring at nuclear fuel cycle facilities could give rise to microbiomes that are characteristic of nuclear activities. These microbiomes could provide valuable information for monitoring nuclear fuel cycle facilities, planning environmental sampling campaigns, and developing biosensor technology for the detection of undisclosed fuel cycle activities and proliferation concerns.

59 BASIC BIOLOGICAL SCIENCES↗

Metagenomes and Metagenome-Assembled Genomes from Microbial Communities in a Biological Nutrient Removal Plant Operated at Hamptons Road Sanitation District (HRSD) with High and Low Dissolved Oxygen Conditions

In this study, we aimed to evaluate Biological Nutrient Removal (BNR) and investigate microbial community changes as the dissolved oxygen is reduced in the aerated portions of wastewater treatment trains. We present a dataset of Metagenome-Assembled Genomes (MAGs) obtained from activated sludge collected from the Hamptons Road Sanitation District (HRSD) BNR plant at the beginning of operation, when the DO was high, and at the end of operation, when the DO was low.

Genomics↗

Metagenomes and Metagenome-Assembled Genomes from Microbial Communities in a Biological Nutrient Removal Plant Operated at Los Angeles County Sanitation District (LACSD) with High and Low Dissolved Oxygen Conditions

In this study, we aimed to evaluate Biological Nutrient Removal (BNR) and investigate microbial community changes as the dissolved oxygen is reduced in the aerated portions of wastewater treatment trains. We present a dataset of Metagenome-Assembled Genomes (MAGs) obtained from activated sludge collected from the Los Angeles County Sanitation District (LACSD) BNR plant at the beginning of operation, when the DO was high, and at the end of operation, when the DO was low.

Genomics↗

Genomic factors limiting the diversity of Saccharomycotina plant pathogens

We compared the genomes of 12 plant-pathogenic Saccharomycotina strains to 360 plant-associated strains to identify features unique to the phytopathogens. Characterization of the oxylipin synthesis genes, a compound believed to be involved in Eremothecium pathogenicity, did not reveal any differences in gene presence within or between the plant-pathogenic and plant-associated strains. A reverse-ecological approach, however, revealed that plant pathogens lack several metabolic enzymes known to assist other phytopathogens in overcoming plant defenses.

phytopathogen↗

A genomic regulatory network for development

Development of the body plan is controlled by large networks of regulatory genes. A gene regulatory network that controls the specification of endoderm and mesoderm in the sea urchin embryo is summarized here. The network was derived from large-scale perturbation analyses, in combination with computational methodologies, genomic data, cis-regulatory analysis, and molecular embryology. The network contains over 40 genes at present, and each node can be directly verified at the DNA sequence level by cis-regulatory analysis. Its architecture reveals specific and general aspects of development, such as how given cells generate their ordained fates in the embryo and why the process moves inexorably forward in developmental time.

Non-NASA Center↗

Radiation induces genomic instability and mammary ductal dysplasia in Atm heterozygous mice

Ataxia-telangiectasia (AT) is a genetic syndrome resulting from the inheritance of two defective copies of the ATM gene that includes among its stigmata radiosensitivity and cancer susceptibility. Epidemiological studies have demonstrated that although women with a single defective copy of ATM (AT heterozygotes) appear clinically normal, they may never the less have an increased relative risk of developing breast cancer. Whether they are at increased risk for radiation-induced breast cancer from medical exposures to ionizing radiation is unknown. We have used a murine model of AT to investigate the effect of a single defective Atm allele, the murine homologue of ATM, on the susceptibility of mammary epithelial cells to radiation-induced transformation. Here we report that mammary epithelial cells from irradiated mice with one copy of Atm truncated in the PI-3 kinase domain were susceptible to radiation-induced genomic instability and generated a 10% incidence of dysplastic mammary ducts when transplanted into syngenic recipients, whereas cells from Atm(+/+) mice were stable and formed only normal ducts. Since radiation-induced ductal dysplasia is a precursor to mammary cancer, the results indicate that AT heterozygosity increases susceptibility to radiogenic breast cancer in this murine model system.

Non-NASA Center↗

A genomic timescale of prokaryote evolution: insights into the origin of methanogenesis, phototrophy, and the colonization of land

BACKGROUND: The timescale of prokaryote evolution has been difficult to reconstruct because of a limited fossil record and complexities associated with molecular clocks and deep divergences. However, the relatively large number of genome sequences currently available has provided a better opportunity to control for potential biases such as horizontal gene transfer and rate differences among lineages. We assembled a data set of sequences from 32 proteins (approximately 7600 amino acids) common to 72 species and estimated phylogenetic relationships and divergence times with a local clock method. RESULTS: Our phylogenetic results support most of the currently recognized higher-level groupings of prokaryotes. Of particular interest is a well-supported group of three major lineages of eubacteria (Actinobacteria, Deinococcus, and Cyanobacteria) that we call Terrabacteria and associate with an early colonization of land. Divergence time estimates for the major groups of eubacteria are between 2.5-3.2 billion years ago (Ga) while those for archaebacteria are mostly between 3.1-4.1 Ga. The time estimates suggest a Hadean origin of life (prior to 4.1 Ga), an early origin of methanogenesis (3.8-4.1 Ga), an origin of anaerobic methanotrophy after 3.1 Ga, an origin of phototrophy prior to 3.2 Ga, an early colonization of land 2.8-3.1 Ga, and an origin of aerobic methanotrophy 2.5-2.8 Ga. CONCLUSIONS: Our early time estimates for methanogenesis support the consideration of methane, in addition to carbon dioxide, as a greenhouse gas responsible for the early warming of the Earths' surface. Our divergence times for the origin of anaerobic methanotrophy are compatible with highly depleted carbon isotopic values found in rocks dated 2.8-2.6 Ga. An early origin of phototrophy is consistent with the earliest bacterial mats and structures identified as stromatolites, but a 2.6 Ga origin of cyanobacteria suggests that those Archean structures, if biologically produced, were made by anoxygenic photosynthesizers. The resistance to desiccation of Terrabacteria and their elaboration of photoprotective compounds suggests that the common ancestor of this group inhabited land. If true, then oxygenic photosynthesis may owe its origin to terrestrial adaptations.

Methane/metabolism↗