Search NASA⌕ Search

SEARCH · Search NASA

Results for “Base Sequence”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 181 records · Page 10

Chemometrics and visible diffuse reflectance spectroscopy to classify plutonium dioxide

Diffuse reflectance (DR) spectra in the Vis-NIR (∼380–1050 nm) region were acquired for a series of PuO 2 samples with a spot size of about 10 × 10 μm. Two batches of six PuO 2 samples, synthesized approximately 7.5 months apart, were prepared using both Pu(III) and Pu(IV) oxalate precursors at three distinct calcination temperatures (450, 650, and 950 °C). This yielded a total of 12 PuO 2 samples and 433 DR spectra. The DR spectrum of PuO 2 contained numerous peaks in the visible region, and characteristic features were identified with respect to calcination temperature and chemistry. A distinct peak multiplet near 615 nm was observed for samples prepared at low calcination temperatures, and a peak near 660 nm was observed for higher calcination temperatures. A multivariate classification strategy based on principal component analysis (PCA) was developed to distinguish PuO 2 calcination temperatures of 450, 650, and 950 °C with 100 % accuracy. Classification results also indicate the potential to distinguish chemical processing history (i.e., Pu(III) or Pu(IV)) based on the spectra with 72 % accuracy based on k-nearest neighbors applied to the PCA scores. Partial least squares discriminant analysis was used to identify variation among batches with 88 % accuracy and found that peaks near 669, 681, 811, and 970 nm were the most useful for predicting the batch identity. Here, this work demonstrates how micro-diffuse reflectance spectroscopy and chemometrics can be used to classify PuO 2 processing history based on Vis-NIR spectral features. Combining the chemometric approach with mapping sequences could provide a rapid, nondestructive approach to classify Pu oxide materials for environmental, forensics, and nonproliferation applications.

Actinide↗

Tuning Copolymer Microstructure Using Ring-Opening Cross-Metathesis Polymerization

The capability of ring-opening cross-metathesis (RO/CM) polymerization to produce alternating copolymers was studied. By treating commercial polybutadiene (PB) with bulky oxanorbornene monomers and Ru-based olefin metathesis catalysts, alternating copolymers were produced under mild conditions with high sequence fidelities. Here, we found that alternating copolymers could be produced starting from a variety of butadiene sources including PB, cyclooctadiene (COD), or t,t,t-1,5,9-cyclododecatriene (CDT), highlighting for the first time the kinetic pathway independence of this process. Kinetic copolymerization analysis of an oxanorbonene monomer with CDT revealed that much higher monomer conversions were obtained compared with the analogous homopolymerizations and showed evidence of alternating monomer incorporation. Copolymerization of these monomers also enabled good control when targeting different molecular weights. Copolymer thermal analysis revealed a strong correlation between thermal behavior and alternating sequence fidelity, providing a second lever beyond composition to tune thermal behavior. These data demonstrate that a broad variety of polymer microstructures can be accessed via RO/CM polymerization and highlight the potential of CDT in alternating copolymer synthesis.

Foster, Jeffrey C. [Oak Ridge National Laboratory ↗

Statistical estimates of the binary properties of rotational variables

ABSTRACT We present a model to estimate the average primary masses, companion mass ranges, the inclination limit for recognizing a rotational variable, and the primary mass spreads for populations of binary stars. The model fits a population’s binary mass function distribution and allows for a probability that some mass functions are incorrectly estimated. Using tests with synthetic data, we assess the model’s sensitivity to each parameter, finding that we are most sensitive to the average primary mass and the minimum companion mass, with less sensitivity to the inclination limit and little to no sensitivity to the primary mass spread. We apply the model to five populations of binary spotted rotational variables identified in ASAS-SN, computing their binary mass functions using RV data from APOGEE. Their average primary mass estimates are consistent with our expectations based on their CMD locations ($\sim 0.75 \, {\rm M}_{\odot }$ for lower main sequence primaries and $\sim 0.9$–$1.2 \, {\rm M}_{\odot }$ for RS CVn and sub-subgiants). Their companion mass range estimates allow companion masses down to $M_2/M_1\simeq 0.1$, although the main sequence population may have a higher minimum mass fraction ($\sim 0.4$). We see weak evidence of an inclination limit $\gtrsim 50^{\circ }$ for the main sequence and sub-subgiant groups and no evidence of an inclination limit in the other groups. No groups show strong evidence for a preferred primary mass spread. We conclude by demonstrating that the approach will provide significantly better estimates of the primary mass and the minimum mass ratio and reasonable sensitivity to the inclination limit with 10 times as many systems.

Phillips, Anya (ORCID:000900051914974X)↗

Purification and expression of a novel bacteriocin, JUQZ-1, against Pseudomonas syringae pv. Actinidiae (PSA), secreted by Brevibacillus laterosporus Wq-1, isolated from the rhizosphere soil of healthy kiwifruit

Kiwifruit canker, caused by Pseudomonas syringae pv. actinidiae (PSA), has led to significant losses in the kiwifruit industry each year. Due to the drug resistance feature of PSA, biological control is currently the most promising method. Developing biocontrol bacteria against PSA could help solve the issue of drug resistance generated during the chemical control of PSA to a certain extent. In this research, a Wq-1 strain that demonstrated excellent inhibitory activity against PSA was isolated from the rhizosphere soil of healthy kiwifruit. Based on the morphological characteristics and phylogenetic analysis of the 16S rRNA gene sequence, the isolated strain was identified as Brevibacillus laterosporus Wq-1. Bacteriostatic proteins were isolated from the cell-free culture filtrate of strain Wq-1 and were found to have a molecular weight of approximately 12 kDa, as determined by sodium dodecyl sulfate-polyacrylamide gel electrophoresis (SDS-PAGE). Liquid chromatography–tandem mass spectrometry (LC–MS/MS) detection revealed that there were several peptides in the target band that were consistent with protein 01021 in the genome. The gene of the 01021 protein was cloned into the plasmid pPICZa, and the recombinant bacteriocin was successfully expressed using the Pichia pastoris X33 expression system. The recombinant protein 01021 effectively inhibited the growth of PSA. This is the first report of the protein’s antimicrobial activity, distinguishing it from previously identified bacteriocins. Therefore, we named this bacteriocin JUQZ-1. In addition, our results showed that the protein JUQZ-1 not only exhibited a broad bacteriostatic spectrum but also high thermal and pH stability suitable for harsh environmental conditions., JUQZ-1, a protein with antimicrobial properties and strong environmental tolerance, may serve as a promising alternative to antibiotics.

Shuai, Yang↗

Tuning the Mechanical Properties of Crosslinked Copolymers via Sequence and Solvent‐Selective Swelling for Vat Photopolymerization

Block copolymers (BCPs) offer distinct advantages for vat photopolymerization by enabling mechanically programmable network structures through microphase-separated morphologies that can be kinetically trapped during curing, yielding properties unattainable in homogeneous resins. However, the respective roles of repeat-unit sequence and solvent environment, together with their interplay in directing network formation and mechanical performance, remain unclear. Here, we synthesize a series of CO 2 -based polycarbonate copolymers comprising a crosslinkable glassy poly(vinyl cyclohexene carbonate) (PVCHC, A block) and a non-crosslinkable soft poly(propylene carbonate) (PPC, B block). The polymer sequence is systematically varied (ABA, BAB, and statistical), and solvent choice controls block-selective swelling to jointly control gelation behavior, microphase morphology, and mechanical response through changes in the accessibility and local environment of photocrosslinkable vinyl groups during network formation, as revealed by photorheology and small angle x-ray scattering. By tuning polymer sequence and curing solvent, we transform nominally identical formulations from brittle to highly ductile materials, achieving a three-orders-of-magnitude range in toughness (0.003 to 9.1 MJ m −3 ). These results establish clear structure–processing–property relationships and identify polymer sequence and selective solvation as powerful strategies for programming both printability and performance of block copolymer resins for additive manufacturing.

additive manufacturing↗

Uncovering heterogeneous intercommunity disease transmission from neutral allele frequency time series

The COVID-19 pandemic has underscored the need for accurate epidemic forecasting to predict pathogen spread, evolution, and evaluate intervention strategies. Forecast reliability hinges on detailed knowledge of disease transmission across population segments, which may be inferred from contact surveys or mobility data. However, these indirect approaches make it difficult to estimate rare transmissions between socially or geographically distant communities. We show that the steep ramp-up of genome sequencing surveillance during the pandemic can be leveraged to directly identify transmission patterns between geographically defined communities. Our approach uses a hidden Markov model to infer the fraction of infections a community imports from others based on how rapidly allele frequencies in the focal community converge to those in the donor communities. Applying this method to SARS-CoV-2 sequencing data from England and the United States, we uncover networks of intercommunity transmission that reflect geographical relationships while exposing significant long-range interactions. The scaling of importation rate with distance is consistent across both countries, yet weaker than expected based on mobility data, highlighting limitations of indirect inference. We show that transmission patterns can change between waves of variants of concern and analyze how the inferred heterogeneity in intercommunity transmission impacts evolutionary forecasts. While applied here to geographically defined communities, our approach could be applied to those defined by other traits (e.g., age, socioeconomic status), provided time-series data can be stratified accordingly. Overall, our study highlights population genomic time series data as a crucial record of epidemiological interactions, which can be deciphered using tree-free inference methods.

Okada, Takashi [Department of Physics; University ↗

MATEY: multiscale adaptive transformer models for spatiotemporal physical systems

Accurate representation of the multiscale features in spatiotemporal physical systems using vision transformer architectures requires extremely long, computationally prohibitive token sequences. To address this issue, we propose two novel adaptive tokenization schemes that dynamically adjust patch sizes based on local features: one ensures convergent behavior to uniform patch refinement, while the other offers better computational efficiency. Moreover, we present a set of spatiotemporal attention schemes, where the temporal or axial spatial dimensions are decoupled, to evaluate their baseline computational and data efficiencies and to determine whether adaptive tokenization can improve this performance. We assess the performance of the proposed multiscale adaptive model, MATEY, in a sequence of experiments. Compared to a full spatiotemporal attention scheme or a scheme that decouples only the temporal dimension, we find that fully decoupled axial attention is less efficient and expressive, requiring more training time and model parameters to achieve the same accuracy. The experiments on the adaptive tokenization schemes show that, compared to a uniformly refined model, the proposed schemes achieve comparable or improved accuracy at a much lower cost in the tested two-dimensional settings. While the asymptotic analysis suggests the potential for favorable scaling, empirical validation at substantially longer sequence lengths remains to be performed in future work. Finally, we demonstrate in two fine-tuning tasks featuring different physics that models pretrained on PDEBench data outperform the ones trained from scratch, especially in the low data regime with frozen attention.

adaptive tokenization↗

Metagenome-assembled genomes from Wind River Basin floodplain sediments Riverton, Wyoming site (May to September 2017)

Microorganisms play a key role in cycling nutrients and contaminants in the terrestrial environment depending on their genetic potential. Here we present metagenome-assembled genomes (MAGs) for the bacterial and archaeal community in floodplain sediment samples taken roughly every month in the period May 18 to September 13 in 2017 at a location (Pit2) close to DOE Legacy Management well 855 at the Riverton, Wyoming floodplain site in the Wind River Basin (WRB). The groundwater at this site exhibits persistent U, Mo, and sulfate plumes and is one of the field sites in focus for the SLAC Groundwater Quality SFA program. Cores were taken with a hand-auger and separated into 5-20 cm segments based on soil horizonation down to 150 cm depth below surface. Each segment was subsampled for microbial analyses. Corresponding 16S rRNA gene amplicon data is available at the NCBI Single Read Archive (SRA) Database BioProject ID PRJNA626616, and soil geochemistry data at doi:10.15485/1631972. 40 metagenomes were sequenced through JGI and can be found under Gold sequencing project: Gs0142591. Metagenomes were assembled, binned, and refined using metawrap to generate MAGs (>50% complete and < 10% contamination based on checkM scores). This dataset includes a zip file of 6993 MAG fasta files and a csv file with quality, taxonomic classification (GTDB RS220), and metagenome accessions for MAGs generated from the Wind River Basin (WRB). This dataset also includes a file-level metadata (flmd.csv) file that lists each file contained in the dataset with associated metadata and a data dictionary (dd.csv) file that contains column/row headers used throughout the files along with a definition, units, and data type.

54 ENVIRONMENTAL SCIENCES↗

An integrated photonic engine for programmable atomic control

Abstract Solutions for scalable, high-performance optical control are important for the development of scaled atom-based quantum technologies. Modulation of many individual optical beams is central to applying arbitrary gate and control sequences on arrays of atoms or atom-like systems. At telecom wavelengths, miniaturization of optical components via photonic integration has pushed the scale and performance of classical and quantum optics far beyond the limitations of bulk devices. However, material platforms for high-speed telecom integrated photonics lack transparency at the short wavelengths required by leading atomic systems. Here, we propose and implement a scalable and reconfigurable photonic control architecture using integrated, visible-light modulators based on thin-film lithium niobate. We combine this system with techniques in free-space optics and holography to demonstrate multi-channel, gigahertz-rate visible beamshaping. When applied to silicon-vacancy artificial atoms, our system enables the spatial and spectral addressing of a dynamically-selectable set of these stochastically-positioned point emitters.

Science & Technology - Other Topics↗

Modeling and Automation Framework for High IBR Integration in Large-Scale Power Systems

The increasing prevalence of power electronics- interfaced renewable generation sources is leading to a gradual replacement of traditional thermal generation-based synchronous machines. In this context, the modeling of a large-scale power grid that incorporates a significant number of inverter-based resources is crucial for understanding the dynamics and effects of these resources on the power system. This study investigates the positive sequence model of grid-following and grid-forming inverters. Additionally, this work explores the integration of distributed energy resources using population as an indicator of their relative geographic locations. To address challenge to integrate these inverter based resources into a realistic grid of the US Western interconnection, automation scripts are developed to streamline the process of replacing conventional generators with grid-following and grid-forming inverters, as well as allocating distributed energy resources. Different penetration levels of these inverters are considered, and their frequency regulation support following a disturbance is compared through dynamic simulations.

Lyu, Xue↗

OrthoPhyl—streamlining large-scale, orthology-based phylogenomic studies of bacteria at broad evolutionary scales

Abstract There are a staggering number of publicly available bacterial genome sequences (at writing, 2.0 million assemblies in NCBI's GenBank alone), and the deposition rate continues to increase. This wealth of data begs for phylogenetic analyses to place these sequences within an evolutionary context. A phylogenetic placement not only aids in taxonomic classification but informs the evolution of novel phenotypes, targets of selection, and horizontal gene transfer. Building trees from multi-gene codon alignments is a laborious task that requires bioinformatic expertise, rigorous curation of orthologs, and heavy computation. Compounding the problem is the lack of tools that can streamline these processes for building trees from large-scale genomic data. Here we present OrthoPhyl, which takes bacterial genome assemblies and reconstructs trees from whole genome codon alignments. The analysis pipeline can analyze an arbitrarily large number of input genomes (>1200 tested here) by identifying a diversity-spanning subset of assemblies and using these genomes to build gene models to infer orthologs in the full dataset. To illustrate the versatility of OrthoPhyl, we show three use cases: E. coli/Shigella, Brucella/Ochrobactrum and the order Rickettsiales. We compare trees generated with OrthoPhyl to trees generated with kSNP3 and GToTree along with published trees using alternative methods. We show that OrthoPhyl trees are consistent with other methods while incorporating more data, allowing for greater numbers of input genomes, and more flexibility of analysis.

59 BASIC BIOLOGICAL SCIENCES↗

Variability in Performance of a Machine Learning Seismicity Catalog: Central Italy, 2016–2017

Machine learning (ML) catalogs contain many more earthquakes than routine catalogs, but their performance in phase picking and earthquake detection has not been fully evaluated. We develop station‐level detection probabilities using logistic regression and combine them across a seismic network to compute spatial magnitude‐of‐completeness fields. We apply this approach to two catalogs from the 2016–2017 Central Italy sequence that were constructed from the same seismic network, one routine and one ML‐based. At the station level, the ML picker increases detection sensitivity by identifying smaller magnitude events and detecting earthquakes at greater distances. Spatially, the magnitude of completeness decreases substantially, with median values shifting from 1.6 to 0.5 for P waves and from 1.7 to 0.5 for S waves. However, the ML catalog also shows greater variability in station‐level performance than the routine catalog. These results demonstrate that ML‐based improvements in detectability are widespread but spatially nonuniform, highlighting their benefits, their limitations, and the potential for further improvements.

15 GEOTHERMAL ENERGY↗

From microbial diversity to functional potential using dimensionality reduction

The high dimensionality of microbial diversity data from ‘omics observations can be reduced using Machine Learning, with many recent studies showcasing ML utility for exploratory ecological feature finding and process prediction. Here, we compare the Self Organizing Map (SOM) dimensionality reduction method to the well-documented sample-based Principal Coordinate Analysis (PCoA) and taxa-based Weighted Gene Correlation Network Analysis (WGCNA) using near daily 16S rRNA gene amplicon sequencing data from the 2019 to 2020 MOSAiC International Arctic Drift Expedition. We then map k-means clustering outputs from each method to available metagenomes, extracting functionally distinct seasonal microbial ecotypes in the surface Arctic Ocean. Our results indicate the SOM method better represented expected seasonal transitions and identified a greater number of metabolically distinct functional groups than the more traditional PCoA ordination. Ultimately, we identified four community ecotypes with distinct taxonomic and functional cut-offs driven by seasonality, water mass, and substrate turnover, highlighting the importance of succession in functional diversity for the central Arctic Ocean. These results reinforce ML dimensionality reduction as a meaningful translator in the mining of historical amplicon datasets to address modern mechanistic questions and potentially provide ’omics informed ecotype diversity to leverage in mechanistic biogeochemical models.

Arctic Ocean↗

Uncertainty quantification and sensitivity analysis of a nuclear thermal propulsion reactor startup sequence

The research presented in this article describes progress in applying stochastic methods, uncertainty quantification, parametric studies, and variance-based sensitivity analysis (also known as Sobol sensitivity analysis) to a full-core model of a nuclear thermal propulsion (NTP) system simulated via the radiation transport code Griffin to simulate neutronics. Our goal is to develop a reduced-order (surrogate) model that can be rapidly sampled with perturbations to multiple input parameters. In this NTP system, reactivity and power feedback affect the rotation of control drums (CDs), which is itself controlled by a hybrid proportional-integral-derivative (PID) controller actuated by the power demand and reactivity feedback from the numerical model. This model uses reactor kinetic feedback (mean generation time [Λ] and effective delayed neutron fraction [ β eff ] from a transient Griffin simulation executed via Griffin’s improved quasi-static solver to provide the kinetic parameters) as inputs to functions that control the CD rotation angle. By investigating numerous stochastic approaches, we developed a dual-purpose surrogate model of the NTP system, using polynomial regression in the Multiphysics Object-Oriented Simulation Environment (MOOSE) Stochastic Tools Module (STM). The trained model can be rapidly sampled while simultaneously perturbing various input parameters, such as coefficients on the PID control or temperature (directly affecting the neutron cross section). The surrogate model delivers accurate (within 5%) results at speeds orders of magnitude faster (minutes, not days of computational time) than the base model. Once the surrogate model has been trained, distributions of the uncertain parameters can be changed at will to investigate the effects of perturbing multiple inputs as well as the effects of these inputs on the model output. For example, coefficients used in the PID control system may vary due to some type of physical interference, or uncertainty may exist in the temperature of the neutron cross sections in various regions of the reactor. A distribution can be placed on these parameters, and operational boundaries can be determined. The goal of this work is to support development of an advanced control system for operating CDs in a functioning NTP system. This work is a scoping study of the MOOSE STM.

21 - SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLAN↗

A copula-based rank histogram ensemble filter

Serial ensemble filters implement triangular probability transport maps to reduce high-dimensional inference problems to sequences of state-by-state univariate inference problems. The univariate inference problems are solved by sampling posterior probability densities obtained by combining constructed prior densities with observational likelihoods according to Bayes' rule. Many serial filters in the literature focus on representing the marginal posterior densities of each state. However, rigorously capturing the conditional dependencies between the different univariate inferences is crucial to correctly sampling multidimensional posteriors. This work proposes a new serial ensemble filter, called the copula rank histogram filter (CoRHF), that seeks to capture the conditional dependency structure between variables via empirical copula estimates; these estimates are used to rigorously implement the triangular (state-by-state univariate) Bayesian inference. The success of the CoRHF is demonstrated on two-dimensional examples and the Lorenz'63 problem. A practical extension to the high-dimensional setting is developed by localizing the empirical copula estimation, and is demonstrated on the Lorenz'96 problem.

97 MATHEMATICS AND COMPUTING↗

DNABERT-S: pioneering species differentiation with species-aware DNA embeddings

SUMMARY: We introduce DNABERT-S, a tailored genome model that develops species-aware embeddings to naturally cluster and segregate DNA sequences of different species in the embedding space. Differentiating species from genomic sequences (i.e. DNA and RNA) is vital yet challenging, since many real-world species remain uncharacterized, lacking known genomes for reference. Embedding-based methods are therefore used to differentiate species in an unsupervised manner. DNABERT-S builds upon a pre-trained genome foundation model named DNABERT-2. To encourage effective embeddings to error-prone long-read DNA sequences, we introduce Manifold Instance Mixup (MI-Mix), a contrastive objective that mixes the hidden representations of DNA sequences at randomly selected layers and trains the model to recognize and differentiate these mixed proportions at the output layer. We further enhance it with the proposed Curriculum Contrastive Learning (C2LR) strategy. Empirical results on 28 diverse datasets show DNABERT-S's effectiveness, especially in realistic label-scarce scenarios. For example, it identifies twice more species from a mixture of unlabeled genomic sequences, doubles the Adjusted Rand Index (ARI) in species clustering, and outperforms the top baseline's performance in 10-shot species classification with just a 2-shot training. AVAILABILITY AND IMPLEMENTATION: Model, codes, and data are publically available at https://github.com/MAGICS-LAB/DNABERT_S.

Zhou, Zhihan↗

LibraryX: A Framework for Cross-Library-Call Optimization

Scientific applications utilize performance libraries as a software engineering concept: these libraries encapsulate important and well-understood (mathematical) operations, allow for reuse, and are implemented and tuned by experts. Domain scientists then implement complex algorithms based on these domainspecific libraries. While individual library calls are optimized, larger performance gains across sequences of calls—sometimes spanning multiple libraries—are often unrealized, forcing a trade-off between performance and implementation complexity.To overcome this issue, we propose LibraryX, an approach and a system that allows for cross-library-call optimization even when library calls stem from multiple performance libraries. LibraryX annotates library calls with semantic information and optimizes entire directed acyclic graphs (DAGs) of calls dynamically using the SPIRAL code generation system. We demonstrate its effectiveness across a range of memory bound workloads, achieving significant speedups on Nvidia, AMD, and Intel accelerators compared to code using native libraries without cross-call optimization.

Rao, Sanil [Carnegie Mellon University,Department ↗

Enabling Grid-Forming Control with Fault Ride-Through in Unbalanced Distribution Networks

Distribution networks are often unbalanced, causing oscillatory responses in inverter control designed for balanced conditions. Here, to address this problem, this paper proposes a novel time-domain transformation appropriate for inverter control and enables the decomposition of three-phase unbalanced signals into constant positive and negative components. Relations useful for calculating unbalanced active and reactive power are derived from first principle, providing insight into vector products of unbalanced three-phase signals. Furthermore, a grid-forming control effective under unbalanced conditions is developed, which delivers superior performance while meeting UNIFI1 specifications for grid-forming control under unbalanced conditions. specifications applicable to category 4 inverter-based resource, like setting and regulating frequency/voltage, providing voltage support, sharing active power, injecting negative sequence current, and riding through faults. A current limiter is proposed for safe fault ride-through and integrates with the grid-forming control featuring frequency/voltage droop controllers and current and voltage control loops. The transformation of interconnected inverters is formulated and stability of the proposed control analyzed to support robust parameter selections. The effectiveness of the proposed transformation and grid-forming control is demonstrated through analytical results and real-time simulation of a IEEE 123 distribution network on the Real-Time Digital Simulator. Comparison with existing methods shows that the proposed strategy satisfies the UNIFI specifications with a much better performance.

24 - POWER TRANSMISSION AND DISTRIBUTION↗