Search NASA⌕ Search

SEARCH · Search NASA

Results for “Sequence Analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 145 records · Page 8

Fuel Property Effects on Stochastic Preignition Events During Engine Load Transitions

Stochastic preignition (SPI) is an abnormal combustion phenomenon that can cause catastrophic engine damage. There have been several proposed mechanisms of SPI, where a uniform source is still not certain, however, SPI tendencies have been shown to be influenced by engine operating conditions, oil composition, engine age, and fuel chemical and physical properties. Laboratory research and testing for SPI propensity is challenging given the stochastic nature of events, as well as the potential for significant degradation of the engine platform and measuring equipment over time. Thus, SPI specific experiments are generally conducted under either sustained or cyclic patterning of steady-state operating conditions to avoid the influence of transient engine boundary conditions on test parameters of interest (e.g. oil additive package, fuel properties, engine speed/load, etc.). In this work a cyclically varying SPI test sequence involves a 5 min engine warmup period at a low engine load of around 4 bar gross indicated mean effective pressure (IMEPg), followed by a transition to high load (~20 bar IMEPg) at a constant 2000 rev/min engine speed for a total of 25 min. This individual test sequence load schedule is then sequentially repeated 10 times to generate significant statistical data for analysis. This work examines the influence of fuel chemical and physical properties on SPI tendency during the unsteady portion of the 10-cycle sequence (the first 5 min of the high load operation in each sequence of the loading cycle) which has been discarded from previous analyses due to the uncertainty in engine operating and thermal boundary conditions. Results from this analysis suggest an increasing trend in the ratio of SPI events during the unsteady test period relative to the steady test period with increasing fuel Reid Vapor Pressure (RVP), implying differences in uncontrolled ignition source terms, possibly from, fuel wall interactions and retention during the load transition phase of the test.

Splitter, Derek [ORNL] (ORCID:0000000174044047)↗

Assembly of small silica nanoparticles using lipid-tethered DNA ‘bonds’

Single-stranded DNA molecules modified with cholesterol functional groups are physically tethered to silica nanoparticles (diameter 25 nm) that are encapsulated in a lipid bilayer. Such tethering increases the azimuthal mobility of the DNA molecules across the nanoparticle surface and enables nonspecific bonding, eliminating the need for specialized surface chemistries (such as silane or thiol ligands). To induce assembly, double-stranded DNA ‘bridge’ molecules are then added with complementary nucleotides to the DNA ‘anchor’ molecules that are physically tethered to the lipids on the surface of the particles. Assembly is observed to occur at room temperature and without the need for temperature annealing. Using automated liquid handling tools, assemblies are created in high throughput and rapidly characterized using SAXS. It is determined that the relative concentration of DNA-to-silica and the ionic strength of the solution are important parameters that affect the resulting assembly. Analysis of SAXS data is performed using coarse-grained particle dynamics simulations. The results support the spontaneous formation of semi-crystalline particle assemblies by particle condensation, where the interparticle distance is tuned by the sequence of the DNA ‘bridge’ used to link the particles. Crystallinity analysis performed on the resulting simulations, optimized to match SAXS observations, suggest that particle clusters display increased crystallinity in the center of the clusters, but their maximum size remains relatively small (sub-micron) before settling occurs, which limits the extent of crystallization.

Chiang, Huat Thart [Univ. of Washington, Seattle, ↗

Impedance Scan of Inverter-Based Resources and Diesel Generator for Stability Analysis: Preprint

Impedance-based methods are widely used for power system stability analysis with inverter-based resources (IBRs), e.g., assessing dynamic interactions between the power grid and an IBR, control interactions between multiple IBRs, and the sub-synchronous oscillation and damping phenomenon. Since it is difficult to get a numerical model 100% matching with the hardware IBR, using the hardware inverter directly to obtain its output impedance has become a prominent approach nowadays. Therefore, this article presents the impedance scan using hardware IBRs, and also a hardware diesel generator as it still stays with the grid before the grid completely goes to renewable. The devices under test (DuTs) for the impedance scan includes two 3-..phi.., 480 V, 60 Hz commercial grid-forming IBRs (one of 250 kVA and another of 125 kVA rating) in series with ..delta..-Y transformers, one 3-..phi.., 480 V, 60 Hz commercial grid-following IBR (of 125 kVA rating), and a 3-..phi.., 480 V, 60 Hz commercial diesel generator (of 187.5 kVA rating). Using voltage signals perturbed with sub-, inter-, and higher harmonic components, and measuring the current response, the positive-sequence impedances are computed via an offline- based post-analysis. Moreover, best-fit transfer functions are estimated that closely resemble the measured data points of the positive-sequence impedances. Based on the observations from various outcomes of the hardware experiments, this article also provides some fundamental insights on the equivalent positive- sequence impedance of a combination of multiple hardware components by comparing the estimated and the empirically computed impedances. A comparative insight on the damping capability of the DuTs using the positive-sequence impedances of the hardware is also discussed.

grid following inverter↗

Impedance Scan of Inverter-Based Resources and Diesel Generator for Stability Analysis

Impedance-based methods are widely used for power system stability analysis with inverter-based resources (IBRs), e.g., assessing dynamic interactions between the power grid and an IBR, control interactions between multiple IBRs, and the sub-synchronous oscillation and damping phenomenon. Since it is difficult to get a numerical model 100% matching with the hardware IBR, using the hardware inverter directly to obtain its output impedance has become a prominent approach nowadays. Therefore, this article presents the impedance scan using hardware IBRs, and also a hardware diesel generator as it still stays with the grid before the grid completely goes to renewable. The devices under test (DuTs) for the impedance scan includes two 3-..phi.., 480 V, 60 Hz commercial grid-forming IBRs (one of 250 kVA and another of 125 kVA rating) in series with ..delta..-Y transformers, one 3-..phi.., 480 V, 60 Hz commercial grid-following IBR (of 125 kVA rating), and a 3-..phi.., 480 V, 60 Hz commercial diesel generator (of 187.5 kVA rating). Using voltage signals perturbed with sub-, inter-, and higher harmonic components, and measuring the current response, the positive-sequence impedances are computed via an offline-based post-analysis. Moreover, best-fit transfer functions are estimated that closely resemble the measured data points of the positive-sequence impedances. Based on the observations from various outcomes of the hardware experiments, this article also provides some fundamental insights on the equivalent positive-sequence impedance of a combination of multiple hardware components by comparing the estimated and the empirically computed impedances. A comparative insight on the damping capability of the DuTs using the positive-sequence impedances of the hardware is also discussed.

current measurement↗

LevSeq: Rapid Generation of Sequence-Function Data for Directed Evolution and Machine Learning

Sequence-function data provides valuable information about the protein functional landscape but is rarely obtained during directed evolution campaigns. Here, we present Long-read every variant Sequencing (LevSeq), a pipeline that combines a dual barcoding strategy with nanopore sequencing to rapidly generate sequence-function data for entire protein-coding genes. LevSeq integrates into existing protein engineering workflows and comes with open-source software for data analysis and visualization. The pipeline facilitates data-driven protein engineering by consolidating sequence-function data to inform directed evolution and provide the requisite data for machine learning-guided protein engineering (MLPE). LevSeq enables quality control of mutagenesis libraries prior to screening, which reduces time and resource costs. Simulation studies demonstrate LevSeq’s ability to accurately detect variants under various experimental conditions. Lastly, we show LevSeq’s utility in engineering protoglobins for new-to-nature chemistry. Widespread adoption of LevSeq and sharing of the data will enhance our understanding of protein sequence-function landscapes and empower data-driven directed evolution.

59 BASIC BIOLOGICAL SCIENCES↗

Methyl formate oxidation kinetics up to 100 atm

Methyl formate (MF, CH3OCHO), the simplest ester, is a representative oxygenated fuel with high oxygen content, and low sooting tendency. However, its oxidation behavior under high-pressure and intermediate-temperature conditions remains insufficiently understood, especially where low-temperature peroxy radical chemistry, methanol chemistry, and pressure-dependent reaction pathways play a critical role. In this study, MF oxidation experiments were conducted in the Princeton supercritical-pressure jet-stirred reactor (SP-JSR) at 20 and 100 atm over the temperature range of 400–950 K under both fuel-lean and fuel-rich conditions. Based on the experimental results, an updated HP-Mech was developed by incorporating previous MF sub-mechanisms, expanded low-temperature peroxy pathways, and evaluated pressure-dependent decomposition kinetics. The newly updated HP-Mech shows greatly improved performance in predicting the onset temperature, the key intermediate species fractions, methanol formation, and the progression of MF oxidation across all the experimental conditions. Path flux analysis indicates that MF consumption at the onset stage is dominated by H-abstraction at the methyl site, forming CH2OCHO radicals that lead to the formation and isomerization of O2CH2OCHO, driving low-temperature chain propagation. Moreover, H-abstraction at the formate site forms CH3OCO radicals that preferentially decompose to CH3, initiating the methanol formation pathway linked to CH3O2 and HO2 chemistry. At the same time, HO2 formation is strongly coupled to MF oxidation through multiple MF-derived radical pathways. HCO originates from MF oxidation and acts as a key coupling species linking fuel consumption to HO2 buildup, especially under high-pressure and intermediate-temperature conditions. In addition to this dominant channel, supplementary HO2 formation pathways involving CH3, CH3O, CH2OH, and CH3O2 reacting with O2 further connect methanol chemistry and oxygenated radical chemistry to the HO2 pool, indicating the central role of HO2 in governing MF oxidation. Sensitivity analysis identifies MF with OH/HO2/CH3O2 reactions and the HO2/H2O2/OH sequence as the key factors controlling reactivity in the high-pressure and intermediate-temperature regime. MF directly reacts with OH/HO2/CH3O2 to consume the fuel and produce reactive radicals like CH2OCHO and CH3OCO that undergo subsequent oxidation pathways. Moreover, HO2 recombination suppresses oxidation at lower temperatures, while thermal decomposition of H2O2 accelerates OH production and promotes fuel consumption as temperature increases. The direct formation of active OH from HO2 radicals further completes the mechanism, improving its prediction especially during the oxidation onset stage.

Low-temperature Chemistry↗

From microbial diversity to functional potential using dimensionality reduction

The high dimensionality of microbial diversity data from ‘omics observations can be reduced using Machine Learning, with many recent studies showcasing ML utility for exploratory ecological feature finding and process prediction. Here, we compare the Self Organizing Map (SOM) dimensionality reduction method to the well-documented sample-based Principal Coordinate Analysis (PCoA) and taxa-based Weighted Gene Correlation Network Analysis (WGCNA) using near daily 16S rRNA gene amplicon sequencing data from the 2019 to 2020 MOSAiC International Arctic Drift Expedition. We then map k-means clustering outputs from each method to available metagenomes, extracting functionally distinct seasonal microbial ecotypes in the surface Arctic Ocean. Our results indicate the SOM method better represented expected seasonal transitions and identified a greater number of metabolically distinct functional groups than the more traditional PCoA ordination. Ultimately, we identified four community ecotypes with distinct taxonomic and functional cut-offs driven by seasonality, water mass, and substrate turnover, highlighting the importance of succession in functional diversity for the central Arctic Ocean. These results reinforce ML dimensionality reduction as a meaningful translator in the mining of historical amplicon datasets to address modern mechanistic questions and potentially provide ’omics informed ecotype diversity to leverage in mechanistic biogeochemical models.

Arctic Ocean↗

Status on Genetic Resistance to Rice Blast Disease in the Post-Genomic Era

Rice blast, caused by Magnaporthe oryzae, is a major threat to global rice production, necessitating the development of resistant cultivars through genetic improvement. Breakthroughs in rice genomics, including the complete genome sequencing of japonica and indica subspecies and the availability of various sequence-based molecular markers, have greatly advanced the genetic analysis of blast resistance. To date, approximately 122 blast-resistance genes have been identified, with 39 of these genes cloned and molecularly characterized. The application of these findings in marker-assisted selection (MAS) has significantly improved rice breeding, allowing for the efficient integration of multiple resistance genes into elite cultivars, enhancing both the durability and spectrum of resistance. Pangenomic studies, along with AI-driven tools like AlphaFold2, RoseTTAFold, and AlphaFold3, have further accelerated the identification and functional characterization of resistance genes, expediting the breeding process. Future rice blast disease management will depend on leveraging these advanced genomic and computational technologies. Emphasis should be placed on enhancing computational tools for the large-scale screening of resistance genes and utilizing gene editing technologies such as CRISPR-Cas9 for functional validation and targeted resistance enhancement and deployment. These approaches will be crucial for advancing rice blast resistance, ensuring food security, and promoting agricultural sustainability.

Pedrozo, Rodrigo↗

Time-series metagenomics reveals changing protistan ecology of a temperate dimictic lake

Abstract Background Protists, single-celled eukaryotic organisms, are critical to food web ecology, contributing to primary productivity and connecting small bacteria and archaea to higher trophic levels. Lake Mendota is a large, eutrophic natural lake that is a Long-Term Ecological Research site and among the world’s best-studied freshwater systems. Metagenomic samples have been collected and shotgun sequenced from Lake Mendota for the last 20 years. Here, we analyze this comprehensive time series to infer changes to the structure and function of the protistan community and to hypothesize about their interactions with bacteria. Results Based on small subunit rRNA genes extracted from the metagenomes and metagenome-assembled genomes of microeukaryotes, we identify shifts in the eukaryotic phytoplankton community over time, which we predict to be a consequence of reduced zooplankton grazing pressures after the invasion of a invasive predator (the spiny water flea) to the lake. The metagenomic data also reveal the presence of the spiny water flea and the zebra mussel, a second invasive species to Lake Mendota, prior to their visual identification during routine monitoring. Furthermore, we use species co-occurrence and co-abundance analysis to connect the protistan community with bacterial taxa. Correlation analysis suggests that protists and bacteria may interact or respond similarly to environmental conditions. Cryptophytes declined in the second decade of the timeseries, while many alveolate groups (e.g., ciliates and dinoflagellates) and diatoms increased in abundance, changes that have implications for food web efficiency in Lake Mendota. Conclusions We demonstrate that metagenomic sequence-based community analysis can complement existing efforts to monitor protists in Lake Mendota based on microscopy-based count surveys. We observed patterns of seasonal abundance in microeukaryotes in Lake Mendota that corroborated expectations from other systems, including high abundance of cryptophytes in winter and diatoms in fall and spring, but with much higher resolution than previous surveys. Our study identified long-term changes in the abundance of eukaryotic microbes and provided context for the known establishment of an invasive species that catalyzes a trophic cascade involving protists. Our findings are important for decoding potential long-term consequences of human interventions, including invasive species introduction.

59 BASIC BIOLOGICAL SCIENCES↗

A multi-omic characterization of the physiological responses to salt stress in Scenedesmus obliquus UTEX393

Scenedesmus obliquus UTEX393 is a promising microalgal candidate for sustainable biomanufacturing but its limited halotolerance hinders large-scale cultivation in saline environments. To investigate the molecular basis of salt stress responses, we conducted a comprehensive multi-omic analysis integrating genomics, transcriptomics, proteomics, lipidomics, metabolomics, and DNA affinity purification sequencing (DAP-seq). An improved nuclear genome assembly and annotation yielded 19,017 gene models and a 97% BUSCO completeness score, enabling construction of a genome-scale metabolic model. Comparing 15 ppt salinity stress to 5 ppt control, growth and productivity were significantly reduced, accompanied by widespread transcriptomic and proteomic changes. Transcriptomic analysis revealed downregulation of photosynthetic machinery and energy conservation genes, and upregulation of stress-responsive elements such as expansins, flavodoxins, and osmoprotectants. Lipidomic profiling showed accumulation of triacylglycerols (TAGs) and degradation of galactosyl lipids, consistent with a shift toward lipid biosynthesis to mitigate redox imbalance. Depletion of key polar metabolites and branched-chain amino acids suggested a rerouting of central carbon metabolism under stress. DAP-seq identified key transcription factors, including LHY1 and SPL12, that target central metabolic enzymes involved in redox balancing, such as glyceraldehyde-3-phosphate dehydrogenase (GAPDH) and malate dehydrogenase (MDH). These findings establish a regulatory-metabolic framework linking redox stress to lipid accumulation and reveal potential engineering targets to enhance salt tolerance. Overall, the multi-omic analysis supports the “overflow” hypothesis, where impaired photosynthesis results in excess reducing equivalents being diverted into TAG synthesis and highlights transcriptional regulators as candidates for improving algal robustness in brackish environments.

09 BIOMASS FUELS↗

Radioisotope Identification with List-Mode Gamma-Ray Data

This work explores the potential of utilizing temporal data from gamma-ray detectors, known as list-mode data, to enhance radioisotope identification. Traditional identification methods, which rely on full gamma-ray spectrum analysis, often require long dwell times and struggle with spectra containing similarly spaced spectral peaks. We hypothesize that by leveraging the probabilistic nature of nuclear decay and the time-encoded information from decay sequences and interactions with surrounding materials, we can improve classification accuracy over static spectral analysis. This research examines the temporal content of list-mode data through exploratory data analysis via correlation discovery and qualitative distribution analysis. Additionally, we propose a probabilistic classification model that can utilize spectral data, temporal data, or both to determine if the incorporation of temporal information improves radioisotope identification. Our findings suggest that the temporal information present in list-mode gamma-ray data has merit and should be further investigated to develop more robust and optimal methods for utilizing this temporal information in applications requiring radioisotope identification.

List-mode data↗

The reference genome for the northeastern Pacific bull kelp, Nereocystis luetkeana

Bull kelp, Nereocystis luetkeana, is a northeastern Pacific kelp with broad distribution from Alaska to central California. Its population declines have caused severe concerns in northern California, the Salish Sea in Washington, and recently in some populations in Oregon. Despite bull kelp's accumulated ecological and physiological studies, an assembled and annotated genomic reference was still unavailable. Here, we report the complete and annotated genome of Nereocystis luetkeana, produced by the California Conservation Genomics Project (CCGP), which aims to reveal genomic diversity patterns across California by sequencing the complete genomes of approximately 150 carefully selected species. The genome was assembled into 1562 scaffolds with 449.82 Mb, 80x of coverage and 22 952 gene models. BUSCO assembly showed a completeness score of 72% for the stramenopiles gene set. The mitochondria and chloroplast genome sequences have 37 Kb and 131 Mb, respectively. The orthology analysis between 10 Phaeophycean genomes showed 1065 expanded and 286 unique orthogroups for this species. Pairwise comparisons showed 542 orthogroups present only in N. luetkeana and M. pyrifera, another large-body kelp. The enrichment analysis of these orthogroups showed important functions related to central metabolism and signaling due to ATPases enrichment in these two species. This genome assembly will provide an essential resource for the ecology, evolution, conservation, and breeding of bull kelp.

California Conservation Genomics Project—CCGP↗

Human Liver Epithelial Cells (HuH7) Response to HCoV-229E Infection Epigenomics (ATAC-Seq) (ACS-DP4)

The purpose of this experiment was to evaluate how wild-type Human coronavirus strain 229E (HCoV-299E) infection alters chromatin accessibility in infected cells. Sample data was obtained from mock-infected cells, UV-inactivated virus treated cells, and replication competent HCoV-229E infected immortalized human liver cells (HuH7) at 24 hours post infection. Samples were processed using ATAC-seq methods for reported bar coded libraries. Sample data was acquired using an Illumina Hi-Seq 2500 sequencer system and further processed for ATAC-Seq expression analysis.

59 BASIC BIOLOGICAL SCIENCES↗

Beyond microbial abundance: metadata integration enhances disease prediction in human microbiome studies

Multiple studies have highlighted the interaction of the human microbiome with physiological systems such as the gut, immune, liver, and skin, via key axes. Advances in sequencing technologies and high-performance computing have enabled the analysis of large-scale metagenomic data, facilitating the use of machine learning to predict disease likelihood from microbiome profiles. However, challenges such as compositionality, high dimensionality, sparsity, and limited sample sizes have hindered the development of actionable models. One strategy to improve these models is by incorporating key metadata from both the human host and sample collection/processing protocols. This remains challenging due to sparsity and inconsistency in metadata annotation and availability. In this paper, we introduce a machine learning-based pipeline for predicting human disease states by integrating host and protocol metadata with microbiome abundance profiles from 68 different studies, processed through a consistent pipeline. Our findings indicate that metadata can enhance machine learning predictions, particularly at higher taxonomic ranks like Kingdom and Phylum, though this effect diminishes at lower ranks. Our study leverages a large collection of microbiome datasets comprising 11,208 samples, therefore enhancing the robustness and statistical confidence of our findings. This work is a critical step toward utilizing microbiome and metadata for predicting diseases such as gastrointestinal infections, diabetes, cancer, and neurological disorders.

Mathematics and Computing↗

Altering translation allows E. coli to overcome chemically stabilized G-quadruplexes

To investigate the effects of stabilizing G-quadruplexes on Escherichia coli, cells were grown in the presence of the G-quadruplex stabilizing compound, NMM, and RNA-seq was performed. This was done for the control cell strain (MG1655 dtolC) as well as a strain with a decrease in a translation elongation factor (MG1655 dtolC tufA::kan). TPM values were used for differential gene expression analysis between each growth condition. All samples were grown and sequenced in triplicate.

Keck, James L.↗

Investigation of the notch sensitivity of tailorable long fiber discontinuous prepreg composite laminates

Tailorable discontinuous fiber composite laminates provide relative formability beyond that of continuous fiber laminates, while achieving improved mechanical performance over comparable stochastic systems. Here, in this work, the notch sensitivity of engineered prepreg platelet molded composite (PPMC) laminates is investigated using the open-hole tension (OHT) test and compared to available data for stochastic PPMCs and continuous fiber laminates made with the same material. The press-formed thermoplastic composites (AS4/PEKK) were molded with a quasi-isotropic stacking sequence. The discontinuous PPMC laminate was found to be notch insensitive with OHT strengths ranging from 145.4 MPa (CV $=$ 7%) for d/w $=$ 0.5 to 229.3 MPa (CV $=$ 9%) for d/w $=$ 0.25. The highly ordered meso-structure of the engineered PPMC laminate yields comparatively excellent mechanical properties for relatively thin laminates in contrast to stochastic systems. Both net- and gross-section failures were observed for d/w $=$ 0.25, which suggests that the engineered PPMC laminates studied here maintain a degree of inherent, internal stress concentrations that compete with those caused by geometric features such as a circular hole. Computational simulations of the OHT tests with explicitly represented platelets were found to be in good agreement with experimental measurements. The progressive failure analysis was used to conduct a numerical investigation of the stacking sequence and platelet meso-morphology.

36 MATERIALS SCIENCE↗

Radioisotope Identification with List-Mode Gamma Ray Data: A rigorous assessment on the value of temporal information applied to radioisotope identification.

This work explores the potential of utilizing temporal data from gamma-ray detectors, known as list-mode data, to enhance radioisotope identification. Traditional identification methods, which rely on full gamma-ray spectrum analysis, often require long dwell times and struggle with “confuser” sources, or spectra with similarly spaced spectral peaks. We hypothesize that by leveraging the probabilistic nature of nuclear decay and the time-encoded information from decay sequences and interactions with surrounding materials, we can improve classification accuracy over static spectral analysis. This research rigorously examines the temporal content of list-mode data through exploratory data analysis via correlation discovery and information theory. We further propose a basic classification model that can utilize spectral or temporal data (or both) to determine if the incorporation of temporal information can improve radioisotope identification. The findings suggest that the temporal information present in list-mode gamma-ray data has merit and should be further investigated.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Sequence Diagrams & PFMEA Table - VGI [SWR-25-107]

As part of the VGI work under the National Charging Experience (ChargeX) Consortium, reliability analysis of communication interfaces for multiple SCM/VGI use-cases was performed using a Process Failure Modes and Analysis (PFMEA) style framework. This repository hosts all the relevant files for each of these use-cases which include: *A visual representation of their communication architecture: Image file (.png) *UML sequence diagram: Plant-UML source file (.puml). Visio file (.vsdx) and image file (.png) derived from the UML sourceX` *The PFMEA table: Excel file (.xlsx) These files are meant to serve as a starting point and can be adapted to company / organization specific SCM implementation.

Gadamsetty, Pranav [National Renewable Energy Labo↗