Search NASASearch

SEARCH · Search NASA

Results for “phylogenetic trees”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

Transcriptomics outputs and phylogenetic trees used for pathway discovery of diterpenoid alkaloids in Delphinium and Aconitum

Transcriptome assemblies, open reading frames in nucleotide and peptide sequences, clustered transcriptomes and corresponding amino acid files, and expression matrices in TPM and raw counts for RNA-seq datasets from Delphinium grandiflorum, Aconitum plicatum, Aconitum lycoctonum, Aconitum carmichaelii, Aconitum japonicum, Aconitum kusnezoffii, and Aconitum vilmorinianum. Also included are phylogenetic trees for terpene synthases and cytochromes P450 mined from these assemblies.

biosynthesis

Sempervirens: A Fast Reconstruction Algorithm for Noisy and Incomplete Binary Matrix Representations of Trees

Applications such as reconstructing cell lineage trees (represented as phylogenetic trees) from single-cell sequencing data require reconstructing a {0,1}-matrix that has many errors and missing entries. We introduce Sempervirens, a very fast matrix reconstruction algorithm for noisy and incomplete matrix representations of phylogenetic trees. Sempervirens uses an iterative maximum-likelihood approach to determine the topology tree represented by the corrupted data. We show that Sempervirens is at least three orders of magnitude faster than other methods on thousand by thousand matrices, with the speed gap widening with larger matrices. We also show that Sempervirens matches state-of-the-art methods in reconstruction accuracy. The speed of Sempervirens enables it to be tractably applied to reconstructing much larger matrices than those that other methods can reconstruct. In addition to experimental results, we justify the algorithm with a mathematical treatment of its subprocedures.

algorithms

Genome collection processing for “Conserved upper thermal limits and small safety margins in soil copiotrophic bacteria”

We extracted the genomic DNA of 400 randomly selected isolates using a Quick-DNA Microprep Kit (Zymo Research D3020) according to the manufacturer’s protocol. We then submitted the extracted gDNA samples for short-read Illumina sequencing (200 Mbp) at SeqCoast Genomics (Portsmouth, NH, USA). After preprocessing the sequences using Trimmommatic (Bolger et al. 2014), we assembled the genomes using SPADES (Bankevich et al. 2012) and checked the quality of each assembly using QUAST (Gurevich et al. 2013). We processed the genome assemblies using a KBase (v1.4.0) pipeline (Allen et al. 2017; Arkin et al. 2018). Briefly, we used DRAM (v0.1.2) with default settings to annotate the genome assemblies. We then evaluated genome quality and possible contamination levels using CheckM (v1.0.18) (Parks et al. 2015) and retained genomes with completeness above 98% and contamination below 5% (n = 354), following the authors' guidelines. We then obtained taxonomic assignments for all remaining isolates using the Genome Taxonomy Database tool GTDB-Tk (v2.3.2, database version r214) (Chaumeil et al. 2019). We constructed a phylogenetic tree using the tool SpeciesTree (v2.2.0). We then trimmed the tree (using Trim SpeciesTree to GenomeSet- v1.4.0), retaining only tips within our collection with measured thermal performance.

59 BASIC BIOLOGICAL SCIENCES

Sugar Relese Supplementary Text and Figures

Phylogenetic tree of GAUT Protein Family and gene model, RNAi construct, and relative transcript abundance of GAUT4 in switchgrass, rice and poplar knockdown (KD) lines.

bio engineered

Small signaling peptides, phylogenetic analysis

ML Phylogenetic analysis of small signaling peptides in Arabidopsis, Sorghum bicolor, Rice, Wheat, Maize, and Brachypodium. Sorghum only phylogenetic trees were used to name genes. All trees but RALF tree are rooted at midpoint.

Kurtz, Evan [Department of Biochemistry and Biophy

Identifying impacts of contact tracing on HIV epidemiological inference from phylogenetic data

Abstract Robust sampling methods are foundational to inferences using phylogenies. Yet the impact of using contact tracing, a type of non-uniform sampling used in public health applications such as infectious disease outbreak investigations, has not been investigated in the molecular epidemiology field. To understand how contact tracing influences a recovered phylogeny, we developed a new simulation tool called SEEPS (Sequence Evolution and Epidemiological Process Simulator) that allows for the simulation of contact tracing and the resulting transmission tree, pathogen phylogeny, and corresponding virus genetic sequences. Importantly, SEEPS takes within-host evolution into account when generating pathogen phylogenies and sequences from transmission histories. Using SEEPS, we demonstrate that contact tracing can significantly impact the structure of the resulting tree, as described by popular tree statistics. Contact tracing generates phylogenies that are less balanced than the underlying transmission process, less representative of the larger epidemiological process, and affects the internal/external branch length ratios that characterize specific epidemiological scenarios. We also examined real data from a 2007–2008 Swedish HIV-1 outbreak and the broader 1998–2010 European HIV-1 epidemic to highlight the differences in contact tracing and expected phylogenies. Aided by SEEPS, we show that the data collection of the Swedish outbreak was strongly influenced by contact tracing even after downsampling, while the broader European Union epidemic showed little evidence of universal contact tracing, agreeing with the known epidemiological information about sampling and spread. Overall, our results highlight the importance of including possible non-uniform sampling schemes when examining phylogenetic trees. For that, SEEPS serves as a useful tool to evaluate such impacts, thereby facilitating better phylogenetic inferences of the characteristics of a disease outbreak. SEEPS is available at https://github.com/MolEvolEpid/SEEPS.

Virology

Ornamental origins and genomic frontiers: a review of big-bracted dogwood research

The big-bracted (Benthamidia) dogwood clade consists of small- to medium-sized deciduous trees within the genus Cornus, known for their showy spring-time floral bract display. Cornus is within the family Cornaceae and order Cornales, and as Cornales is one of the earliest diverging asterids, these taxa have been important for phylogenetic research. Three species within the big-bracted clade, flowering (Cornus florida), kousa (C. kousa), and Pacific (C. nuttallii) dogwoods, are popular ornamental landscape plants in North America, with more than 130 cultivars released. Despite their commercial popularity, numerous research gaps have limited the expansion of fundamental research and dogwood breeding programs. In this present review, we aim to provide a thorough overview of our current understanding of 1) the phylogenetic and biogeographic context, 2) plant biology and major pests and pathogens impacting commercialization, 3) historical commercialization and propagation methods, and 4) genetic and genomic resources and how they have been implemented to understand these species. Research gaps and future directions to advance basic research and breeding of big-bracted ornamental dogwoods are discussed throughout.

Cornus florida

Beyond Solanaceae: incorporation of feruloyltyramine and feruloyloctopamine into Cannabaceae lignins

The ferulic acid amides, feruloyltyramine and feruloyloctopamine, have been widely reported as integral constituents in the lignins in several species of Solanaceae in which they function as authentic lignin monomers. In the present study, we demonstrate that these ferulic acid amides are likewise incorporated into the lignins of species within Cannabaceae, including hemp (Cannabis sativa), hops (Humulus lupulus), and European nettle tree (Celtis australis). Structural analyses using derivatization followed by reductive cleavage (DFRC) and two-dimensional nuclear magnetic resonance (2D-NMR) spectroscopy revealed that these ferulic acid amides are incorporated via 4−O- and 8−O-ether linkages, as well as through 8−5′ linkages forming phenylcoumaran structures. Examination of a broad phylogenetic range of plant families demonstrated the absence of these ferulic acid amides from the lignins of all families studied except Solanaceae and Cannabaceae. Given the distant phylogenetic relationship between Solanaceae and Cannabaceae, the recruitment of these ferulic acid amides as lignin monomers in both lineages likely constitutes a case for convergent evolution at the level of lignin biosynthetic pathways. The significance of these ferulic acid amides lies in their unique role as the sole nitrogen-containing phenolic compounds known to participate in lignin formation.

Cannabaceae

Diversity of Sordariales Fungi: Identification of Seven New Species of Naviculisporaceae Through Morphological Analyses and Genome Sequencing

Thanks to next-generation sequencing (NGS) technologies, the diversity of fungi can now be investigated through the analysis of their genome sequences. Naviculisporaceae is a family within the Sordariales, whose diversity is not well-known, with only one genome sequence published for this family. Here, we report on the isolation and cultivation of 20 new strains of Naviculisporaceae. Their genome sequences, as well as those of the five commercially available strains, were determined, thus providing complete genome sequences for 25 new Naviculisporaceae strains. Species delimitation was conducted using a combination of (1) ITS + LSU phylogenetic analysis of the new isolates along with other known species of the family, (2) comparisons between DNA barcode sequences of the new strains with those of the known species, and (3) average genome-wide nucleotide identity calculation. We built a phylogenomic tree and studied the organization of the mating-type locus. In vitro fruiting was obtained for 16 strains, enabling the definition of seven new species, namely Pseudorhypophila gallica, Pseudorhypophila guyanensis Rhypophila alpibus, Rhypophila brasiliensis, Rhypophila camarguensis, Rhypophila reunionensis and Rhypophila thailandica, as well as two new combinations, namely Pseudorhypophila latipes and Pseudorhypophila oryzae. Eight strains for which in vitro fruiting was not obtained may belong to additional new species. These results expand the known diversity of the Naviculisporaceae and greatly enlarge the genomic data available for the family.

Naviculisporaceae

A haplotype‐resolved reference genome of Quercus alba sheds light on the evolutionary history of oaks

Summary White oak ( Quercus alba ) is an abundant forest tree species across eastern North America that is ecologically, culturally, and economically important. We report the first haplotype‐resolved chromosome‐scale genome assembly of Q. alba and conduct comparative analyses of genome structure and gene content against other published Fagaceae genomes. We investigate the genetic diversity of this widespread species and the phylogenetic relationships among oaks using whole genome data. Despite strongly conserved chromosome synteny and genome size across Quercus , certain gene families have undergone rapid changes in size, including defense genes. Unbiased annotation of resistance (R) genes across oaks revealed that the overall number of R genes is similar across species – as are the chromosomal locations of R gene clusters – but, gene number within clusters is more labile. We found that Q. alba has high genetic diversity, much of which predates its divergence from other oaks and likely impacts divergence time estimations. Our phylogenetic results highlight widespread phylogenetic discordance across the genus. The white oak genome represents a major new resource for studying genome diversity and evolution in Quercus . Additionally, we show that unbiased gene annotation is key to accurately assessing R gene evolution in Quercus .

Larson, Drew A. [Department of Biology Indiana Uni

Phylogenomic insights into the taxonomy, ecology, and mating systems of the lorchel family Discinaceae (Pezizales, Ascomycota)

Lorchels, also known as false morels (Gyromitra sensu lato), are iconic due to their brain-shaped mushrooms and production of gyromitrin, a deadly mycotoxin. Molecular phylogenetic studies have hitherto failed to resolve deep-branching relationships in the lorchel family, Discinaceae, hampering our ability to settle longstanding taxonomic debates and to reconstruct the evolution of toxin production. We generated 75 draft genomes from cultures and ascomata (some collected as early as 1960), conducted phylogenomic analyses using 1542 single-copy orthologs to infer the early evolutionary history of lorchels, and identified genomic signatures of trophic mode and mating-type loci to better understand lorchel ecology and reproductive biology. Our phylogenomic tree was supported by high gene tree concordance, facilitating taxonomic revisions in Discinaceae. We recognized 10 genera across two tribes: tribe Discineae (Discina, Maublancomyces, Neogyromitra, Piscidiscina, and Pseudodiscina) and tribe Gyromitreae (Gyromitra, Hydnotrya, Paragyromitra, Pseudorhizina, and Pseudoverpa); Piscidiscina was newly erected and 26 new combinations were formalized. Paradiscina melaleuca and Marcelleina donadinii formed their own family-level clade sister to Morchellaceae, which merits further taxonomic study. Genome size and CAZyme content were consistent with a mycorrhizal lifestyle for the truffle species (Hydnotrya spp.), whereas the other Discinaceae genera possessed genomic properties of a saprotrophic habit. Lorchels were found to be predominantly heterothallic-either MAT1-1 or MAT1-2-but a single occurrence of colocalized mating-type idiomorphs indicative of homothallism was observed in Gyromitra esculenta strain CBS101906 and requires additional confirmation and follow-up study. Lastly, we confirmed that gyromitrin has a phylogenetically discontinuous distribution, having been detected exclusively in two distantly related genera (Gyromitra and Piscidiscina) belonging to separate tribes. Our genomic dataset will facilitate further investigations into the gyromitrin biosynthesis genes and their evolutionary history. With additional sampling of Geomoriaceae and Helvellaceae-two closely related families with no publicly available genomes-these data will enable comprehensive studies on the independent evolution of truffles and ecological diversification in an economically important group of pezizalean fungi.

Dirks, Alden C

A haplotype-resolved, chromosome-scale genome assembly for the southern live oak, Quercus virginiana

Hybridization is a major force driving diversification, migration, and adaptation in Quercus species. While population genetics and phylogenetics have traditionally been used for studying these processes, advances in sequencing technology now enable us to incorporate comparative and pan-genomic approaches as well. Here, we present a highly contiguous, chromosome-scale and haplotype-resolved genome assembly for the southern live oak, Quercus virginiana, the first reference genome for section Virentes, as part of the American Campus Tree Genomes program. Originating from a clone of Auburn University's historic “Toomer's Oak,” this assembly contributes to the pool of genomic resources for investigating recombination, haplotype variation, and structural genomic changes influencing hybridization potential in this clade and across Quercus. It also provides insights into the architecture of the putative centromeric regions within the genus. Alongside other oak references, the Q. virginiana genome will support research into the evolution and adaptation of the Quercus genus.

Quercus virginiana

Convergent expansions of keystone gene families drive metabolic innovation in Saccharomycotina yeasts

Many remarkable phenotypes have repeatedly occurred across vast evolutionary distances. When convergent traits emerge on the tree of life, they are sometimes driven by the same underlying gene families, while other times, many different gene families are involved. Conversely, a gene family may be repeatedly recruited for a single trait or many different traits. To understand the general rules governing convergence at both genomic and phenotypic levels, we systematically tested associations between 56 binary metabolic traits and gene count in 14,785 gene families from 993 Saccharomycotina yeasts. Using a recently developed phylogenetic approach that reduces spurious correlations, we found that gene family expansion and contraction were significantly linked to trait gain and loss in 45/56 (80%) traits. While 595/739 (81%) significant gene families were associated with only one trait, we also identified several “keystone” gene families that were significantly associated with up to 13/56 (23%) of all traits. Strikingly, most of these families are known to encode metabolic enzymes and transporters, including all members of the industrially relevant MAL tose fermentation loci in the baker’s yeast Saccharomyces cerevisiae. These results indicate that convergent evolution on the gene family level may be more widespread across deeper timescales than previously believed.

59 BASIC BIOLOGICAL SCIENCES

Divergent trait controls on soluble sugars and starch underlie global strategies of tree carbohydrate storage

Nonstructural carbohydrate (NSC) stores buffer tree metabolism, osmotic regulation, and defense, thereby mediating tolerance and survival under climate extremes. Yet, the functional and evolutionary determinants of interspecific variation in NSC remain elusive, limiting understanding and prediction of forest carbon allocation and mortality under global change. Here, we present a cross-species synthesis of NSC concentrations across multiple organs for 281 woody species from 102 mixed forest communities worldwide, where we quantified species-specific deviations from community means to disentangle intrinsic trait effects from environmental and methodological variation. We found phylogenetic signals in NSC deviations, with coniferous gymnosperms and evergreen species consistently maintaining lower stem soluble sugars and starch concentrations than co-occurring angiosperms and deciduous species, respectively. A global pattern emerged where greater stomatal sensitivity to leaf water potential was associated with declines in the relative concentrations of both sugars and starch. In contrast, xylem hydraulic safety traits showed weak and organ-dependent relationships with NSC concentrations. Sugars increased with photosynthetic capacity and declined with wood density, whereas starch showed the reverse pattern, which aligned with the distinct functional-metabolic roles of sugars and starch. By integrating trait-based ecology with a community-centered framework, our study provides global evidence that stomatal regulation, photosynthetic capacity, specific leaf area, and wood density jointly govern interspecific NSC variation, through contrasting effects on sugars and starch. These are among the most broadly measured traits globally, thus the emergent carbohydrate–trait relationships can have broad applications toward understanding and predicting forest growth and survival under climate change.

tropic system

Global metagenomics reveals plastid diversity and unexplored algal lineages

Photosynthetic organelles in eukaryotes originated through primary endosymbiosis with a cyanobacterium, an event that profoundly shaped the evolutionary landscape of the eukaryotic tree of life. Primary plastids in Archaeplastida, especially in cultivable plants and algae, contribute most to known plastid diversity. Secondary and higher-order endosymbiosis, involving eukaryotic hosts and algal endosymbionts, further spread photosynthesis among protists within the CASH lineages (Cryptophyta, Alveolata, Stramenopila, and Haptophyta). Despite various hypotheses explaining secondary plastid evolution and distribution, empirical support remains limited. Here, we employ cultivation-independent global metagenomics to expand plastid diversity and investigate plastid origins. We capture 1,027 plastid sequences, including 300 novel sequences belonging to previously unsequenced plastids and representing yet-to-be described microeukaryotes. This includes a new lineage that offers insights into plastid evolution in haptophytes and cryptophytes. Our results confirm that Archaeplastida plastids originate from an early branching cyanobacterial lineage closely related to Gloeomargaritales and identify the closest extant relative of Paulinella plastids. Additionally, our findings suggest two independent origins of secondary red-algal plastids, contributing to plastid diversity in CASH lineages and challenging the prevailing model of single secondary plastid origin. Our study highlights the importance of metagenomic data in uncovering biological diversity and advancing understanding of plastid relationships across photosynthetic eukaryotes.

59 BASIC BIOLOGICAL SCIENCES