Search NASA⌕ Search

SEARCH · Search NASA

Results for “cluster analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 523 records · Page 29

Comparative techniques used to evaluate Thematic Mapper data for land cover classification in Logan County, West Virginia

Several digital data processing techniques were evaluated in an effort to identify and map active/abandoned, partially reclaimed, and fully revegetated surface mine areas in the central portion of Logan County. The TM data were first subjected to various enhancement procedures, including a linear contrast stretch, principal components and canonical analysis transformations. At the same time, four general procedures were followed to produce six classifications as a means of comparing the techniques involved. Preliminary results show that various feature extraction/data reduction techniques provide classification results equal or superior to the more straightforward unsupervised clustering technique. Analyst interaction time for labelling clusters is reduced using the canonical analysis and principal components procedures, though the canonical technique has clearly produced better results to date.

Brumfield, J. O.↗

The S-PLUS Fifth Data Release: Over 4500 Square Degrees of the Southern Sky and a Multicolor View of the Hydra and Antlia Galaxy Clusters

Abstract We present the fifth data release (DR5) of the Southern Photometric Local Universe Survey (S-PLUS), covering 4592 deg 2 across 2491 fields. Observations were conducted with the T80-South, a Brazilian robotic telescope equipped with the Javalambre 12-filter system, containing five broadband and seven narrowband filters. Data products feature FITS images and extensive catalogs containing fluxes, magnitudes, and shape parameters for over 113 million detections. In addition, several value-added catalogs are provided, offering photometric redshifts (photo- zs ), object classifications, masks, and extinction coefficients. For the first time, this release includes coverage of 110 deg 2 along the Galactic disk, facilitating new research into Galactic structure and stellar populations. The release also provides full coverage of the Hydra Supercluster and numerous other nearby clusters with improved data reduction and calibration, enhancing photo- z accuracy, which is vital for large-scale structure studies. A preliminary analysis of the Hydra and Antlia galaxy clusters up to 5 × R 200 yields an updated catalog of 1706 cluster members based on both spectroscopic and high-quality photo- zs . Our photometric data shows that both clusters have a similar proportion of galaxies with an H α excess relative to their clustercentric distance, though Hydra has a higher fraction near its center. Additionally, the spatial distribution of all objects in our sample highlights a bridge connecting both clusters. We verify that S-PLUS DR5 provides a solid foundation for future scientific investigations, ranging from solar system studies to cosmology.

Vinicius Rodrigues de Lima, Erik [Universidade de ↗

Cu site differentiation in tetracopper(I) sulfide clusters enables biomimetic N 2 O reduction

Copper clusters feature prominently in both metalloenzymes and synthetic nanoclusters that mediate catalytic redox transformations of gaseous small molecules. Such reactions are critical to biological energy conversion and are expected to be crucial parts of renewable energy economies. However, the precise roles of individual metal atoms within clusters are difficult to elucidate, particularly for cluster systems that are dynamic under operating conditions. Here, we present a metal site-specific analysis of synthetic Cu 4 (μ 4 -S) clusters that mimic the Cu Z active site of the nitrous oxide reductase enzyme. Leveraging the ability to obtain structural snapshots of both inactive and active forms of the synthetic model system, we analyzed both states using resonant X-ray diffraction anomalous fine structure (DAFS), a technique that enables X-ray absorption profiles of individual metal sites within a cluster to be extracted independently. Using DAFS, we found that a change in cluster geometry between the inactive and active states is correlated to Cu site differentiation that is presumably required for efficient activation of N 2 O gas. More precisely, we hypothesize that the Cu δ+ ∙∙∙Cu δ- pairs produced upon site differentiation are poised for N 2 O activation, as supported by computational modeling. These results provide an unprecedented level of detail on the roles of individual metal sites within the synthetic cluster system and how those roles interplay with cluster geometry to impact the reactivity function. We expect this fundamental knowledge to inform understanding of metal clusters in settings ranging from (bio)molecular to nanocluster to extended solid systems involved in energy conversion.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Radio emission in the directions of cD and related galaxies in poor clusters. III - VLA observations at 20 cm

VLA radio maps and optical identifications of a sample of sources in the directions of 21 Yerkes poor cluster fields are presented. The majority of the cluster radio sources are associated with the dominant D or cD galaxies (approximately 70 percent). Our analysis of dominant galaxies in rich and poor clusters indicates that these giant galaxies are much more often radio emitters (approximately 25 percent of cD's are radio active in the poor clusters), have steeper radio spectra, and have simpler radio morphologies (i.e., double or other linear structure) than other less bright ellipticals. A strong continuum of radio properties in cD galaxies is seen from rich to poor clusters. It is speculated that the location of these dominant galaxies at the cluster centers (i.e., at the bottom of a deep, isolated gravitational potential well) is the crucial factor in explaining their multifrequency activity. Galaxy cannibalism and gas infall models as fueling mechanisms for the observed radio and X-ray emission are discussed

Burns, J. O.↗

Streamlining heterologous expression of top carbonic anhydrases in Escherichia coli : bioinformatic and experimental approaches

Carbonic anhydrase (CA) enzymes facilitate the reversible hydration of CO 2 to bicarbonate ions and protons. Identifying efficient and robust CAs and expressing them in model host cells, such as Escherichia coli, enables more efficient engineering of these enzymes for industrial CO 2 capture. However, expression of CAs in E. coli is challenging due to the possible formation of insoluble protein aggregates, or inclusion bodies. This makes the production of soluble and active CA protein a prerequisite for downstream applications. In this study, we streamlined the process of CA expression by selecting seven top CA candidates and used two bioinformatic tools to predict their solubility for expression in E. coli. The prediction results place these enzymes in two categories: low and high solubility. Our expression of high solubility score CAs (namely CA5-SspCA, CA6-SazCAtrunc, CA7-PabCA and CA8-PhoCA) led to significantly higher protein yields (5 to 75 mg purified protein per liter) in flask cultures, indicating a strong correlation between the solubility prediction score and protein expression yields. Furthermore, phylogenetic tree analysis demonstrated CA class-specific clustering patterns for protein solubility and production yields. Unexpectedly, we also found that the unique N-terminal, 11-amino acid segment found after the signal sequence (not present in its homologs), was essential for CA6-SazCA activity. Overall, this work demonstrated that protein solubility prediction, phylogenetic tree analysis, and experimental validation are potent tools for identifying top CA candidates and then producing soluble, active forms of these enzymes in E. coli. The comprehensive approaches we report here should be extendable to the expression of other heterogeneous proteins in E. coli.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

The dynamical state of eROSITA clusters and its impact on the brightest cluster galaxy luminosity

The first Spectrum-Roentgen-Gamma (SRG) eROSITA public release contains 12 247 clusters and groups from its first 6 months of operation. We used the offset between the brightest cluster galaxy (BCG) and the X-ray peak ( D BCG − X ) to classify the cluster dynamical state of 3946 galaxy clusters and groups. We aim to investigate the evolution of the merger and relaxed cluster distributions with redshift and mass, and the distributions’ impact on the BCG. We used the X-ray peak from the eROSITA survey and the BCG position from the LS DR10 optical data, which includes the DECam eROSITA Survey optical data, to measure the D BCG − X offset. We modelled the distribution of D BCG − X , in units of R 500 , as the sum of two Rayleigh distributions representing the cluster’s relaxed and disturbed populations, and explored their evolution with redshift and mass. To explore the impact of the cluster’s dynamical state on the BCG luminosity, we separated the main sample according to the dynamical state. We defined clusters as relaxed if D BCG − X < 0.25R500, disturbed if D BCG − X > 0.5R500, and ‘diverse’ if 0.25R500 < D BCG − X < 0.5R500. We find no evolution of the merging fraction with redshift or mass. The width of the relaxed distribution increases with redshift, while the width of the two Rayleigh distributions decreases with mass. The analysis reveals that BCGs in relaxed clusters are brighter than BCGs in both the disturbed and diverse cluster populations. The most significant differences are found for high-mass clusters at higher redshifts. The results suggest that BCGs in low-mass clusters are less centrally bound than those in high-mass systems, irrespective of the dynamical state. Over time, BCGs in relaxed clusters progressively align with the potential centre. This alignment correlates with their luminosity growth relative to BCGs in dynamically disturbed clusters, underscoring the critical role of the cluster’s dynamical state in regulating BCG evolution.

galaxy clusters↗

Statistical relationships across epigenomes using large-scale hierarchical clustering

Recent advances in genomics and sequencing platforms have revolutionized our ability to create immense data sets, particularly for studying epigenetic regulation of gene expression. However, the avalanche of epigenomic data is difficult to parse for biological interpretation given nonlinear complex patterns and relationships. This attractive challenge in epigenomic data lends itself to machine learning for discerning infectivity and susceptibility. In this study, we explore over 3000 epigenomes of uninfected individuals and provide a framework to characterize the relationships among epigenetic modifiers, their modifiers, genetic loci, and specific immune cell types across all chromosomes using hierarchical clustering. Hierarchical clustering of epigenomic data revealed consistent epigenetic patterns across chromosomes, demonstrating that variation due to epigenetic modifiers is greater than variation between cell types. Gene Ontology and KEGG pathway analyses indicated significant enrichment of genes involved in chromatin remodeling, mRNA splicing, immune responses, and the regulation of microRNAs and snoRNAs. Epigenetic modifiers frequently formed biologically relevant clusters, including the cohesin complex, RNA Polymerase II transcription factors, and PRC2 complex members. These clustering behaviors remained consistent across all chromosomes, supported by entropy analysis and high Adjusted Rand Index scores, indicating robust cross-chromosomal similarity. Co-occurrence analysis further revealed specific sets of modifiers that consistently appeared together within clusters, reflecting shared biological functions and interactions. Validation using another dataset confirmed the reproducibility of these clustering patterns and modifier co-occurrence relationships, underscoring the reliability and generalizability of the methodology.

97 MATHEMATICS AND COMPUTING↗

X-ray-imaging observations of clusters of galaxies

Einstein X-ray imaging observations, made to illustrate the variety of phenomena that can be considered through X-ray image analysis, are presented. Attention is given to general cluster properties and intracluster gas. Individual clusters are discussed (considering classification and dynamical evolution), and X-ray images are used to determine cluster mass distribution and to examine distant clusters. X-ray observations have contributed information in regard to processes affecting galaxies, the intracluster medium, and the cluster itself. Analyses have traced massive halos around dominant galaxies in unevolved clusters, and have helped define the cluster gravitational potential. In addition, multi-component double clusters have been discovered, and material which has been ram-pressure stripped from a hot corona around the M86 galaxy in Virgo was observed. Finally, quantitative estimates of the fractions of young and evolved clusters and determinations of total cluster mass are possible using X-ray observations.

Forman, W.↗

Dark energy survey: Modeling strategy for multiprobe cluster cosmology and validation for the full six-year dataset

Here, we introduce an updated To&Krause2021 model for joint analyses of cluster abundances and large-scale two-point correlations of weak lensing and galaxy and cluster clustering (termed CL+3×2 pt analysis) and validate that this model meets the systematic accuracy requirements of analyses with the statistical precision of the final Dark Energy Survey (DES) Year 6 (Y6) dataset. The validation program consists of two distinct approaches, (i) identification of modeling and parametrization choices and impact studies using simulated analyses with each possible model misspecification and (ii) end-to-end validation using mock catalogs from customized Cardinal simulations that incorporate realistic galaxy populations and DES-Y6-specific galaxy and cluster selection and photometric redshift modeling, which are the key observational systematics. In combination, these validation tests indicate that the model presented here meets the accuracy requirements of DES-Y6 for CL+3×2 pt based on a large list of tests for known systematics. In addition, we also validate that the model is sufficient for several other data combinations: the CL+GC subset of this data vector (excluding galaxy–galaxy lensing and cosmic shear two-point statistics) and the CL+3×2 pt+BAO+SN (combination of CL+3×2 pt with the previously published Y6 DES baryonic acoustic oscillation and Y5 supernovae data).

79 ASTRONOMY AND ASTROPHYSICS↗

Computer-aided analysis of Landsat-1 MSS data - A comparison of three approaches, including a 'modified clustering' approach

Three approaches for analyzing Landsat-1 data from Ludwig Mountain in the San Juan Mountain range in Colorado are considered. In the 'supervised' approach the analyst selects areas of known spectral cover types and specifies these to the computer as training fields. Statistics are obtained for each cover type category and the data are classified. Such classifications are called 'supervised' because the analyst has defined specific areas of known cover types. The second approach uses a clustering algorithm which divides the entire training area into a number of spectrally distinct classes. Because the analyst need not define particular portions of the data for use but has only to specify the number of spectral classes into which the data is to be divided, this classification is called 'nonsupervised'. A hybrid method which selects training areas of known cover type but then uses the clustering algorithm to refine the data into a number of unimodal spectral classes is called the 'modified-supervised' approach.

Fleming, M. D.↗

Cooling Flow Spectra in Ginga Galaxy Clusters

The primary focus of this research project has been a joint analysis of Ginga LAC and Einstein SSS X-ray spectra of the hot gas in galaxy clusters with cooling flows is reported. We studied four clusters (A496, A1795, A2142 & A2199) and found their central temperatures to be cooler than in the exterior, which is expected from their having cooling flows. More interestingly, we found central metal abundance enhancements in two of the clusters, A496 and A2142. We have been assessing whether the abundance gradients (or lack thereof) in intracluster gas is correlated with galaxy morphological gradients in the host clusters. In rich, dense galaxy clusters, elliptical and SO galaxies are generally found in the cluster cores, while spiral galaxies are found in the outskirts. If the metals observed in clusters came from proto-ellipticals and proto-S0s blowing winds, then the metal distribution in intracluster gas may still reflect the distribution of their former host galaxies. In a research project which was inspired by the success of the Ginga LAC/Einstein SSS work, we analyzed X-ray spectra from the HEAO-A2 MED and the Einstein SSS to look for temperature gradients in cluster gas. The HEAO-A2 MED was also a non-imaging detector with a large field of view compared to the SSS, so we used the differing fields of view of the two instruments to extract spatial information. We found some evidence of cool gas in the outskirts of clusters, which may indicate that the nominally isothermal mass density distributions in these clusters are steepening in the outer parts of these clusters.

White, Raymond E., III↗

Direct Observation of Elusive (DTBM‐SEGPHOS)CuH Monomer Enables Mechanistic Insights Into Hydrocupration, Aggregation, and Dynamics of Alkene Functionalization Catalysis

The bulky diphosphine DTBM-SEGPHOS is widely employed in CuH-catalyzed transformations as it provides remarkably active catalyst systems. The transient (DTBM-SEGPHOS)CuH monomer (LCuH) is the often-invoked active species. However, its instability has prevented spectroscopic characterization and mechanistic elucidation, hindering mechanistic understanding. We report low-temperature NMR spectroscopic characterization of LCuH, enabling quantitative kinetic analysis of the stoichiometric hydrocupration and catalytic hydroboration of cyclopentene, as well as the structural identification of two CuH clusters. LCuH inserts cyclopentene at −43°C, reaffirming its high reactivity toward olefins. LCuH deactivates to form L 2 Cu 3 H 3 and L 2 Cu 4 H 4 clusters, in which LCuH dimerization initiates aggregation. Kinetic analysis of reactions of unactivated alkenes indicates that competing on-cycle alkene hydrocupration and LCuH dimerization impact performance, as catalyst deactivation and turnover occur on comparable timescales. Structure–activity analysis using atomistic simulations shows that the steric profile of DTBM-SEGPHOS increases the CuH dimerization barrier by ∼7.7 kcal mol−1 compared to that of SEGPHOS, rationalizing the unique ability of DTBM-SEGPHOS to stabilize a reactive monomer for hydrocupration of broader alkene substrates. These findings illustrate the fundamental design principle that steric control of aggregation governs CuH catalyst performance, explaining both the exceptional activity of (DTBM-SEGPHOS)CuH and the limitations imposed by competing deactivation.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

STM/S Grid LDOS Data and Analysis Code for Deciphering Majorana Zero Modes in Topological Superconductor

This dataset provides raw millikelvin scanning tunneling microscopy/spectroscopy (STM/S) grid spectroscopy data and Python analysis scripts supporting the manuscript “Deciphering Majorana Zero Modes in Topological Superconductor FeTe0.55Se0.45 with Machine-Learning-Assisted Spectral Deconvolution.” The dataset includes a raw grid spectroscopy file acquired on FeTe0.55Se0.45 at 40 mK under magnetic field, together with Python/Jupytext analysis scripts used for STM/S data processing, visualization, spectral deconvolution, Lorentzian peak fitting, feature extraction, machine-learning-assisted clustering, and figure generation. These files support the analysis of vortex-core local density of states and the identification of zero-bias-peak-related spectral components from complex in-gap states. The dataset is intended to provide a citable archival record of the data and analysis code associated with the published manuscript and to support transparency and reproducibility of the reported STM/S and machine-learning workflow.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Data and scripts associated with “Moisture content modulates DOM thermodynamic regulation of oxygen consumption in drying streambed sediments”

This data package is associated with the publication “Moisture content modulates DOM thermodynamic regulation of oxygen consumption in drying streambed sediments” published in Scientific Reports (Garayburu-Caruso et al., 2026). The package contains processed data products and scripts used to quantify how drying and re-inundation of riverbed sediments influence dissolved organic matter (DOM) thermodynamic properties and their relationship with sediment oxygen (O₂) consumption across 33 stream sites in the contiguous United States. The data package contains DOM thermodynamic metrics (e.g., Gibbs free energy of carbon oxidation and thermodynamic efficiency), and O₂ consumption along with watershed-scale climate and land-cover metrics used as explanatory variables in the analyses. Underlying unprocessed and processed ultrahigh-resolution mass spectrometry data, oxygen consumption rates from laboratory moisture-manipulation experiments, within-sample environmental properties, sediment moisture content and contextual field measurements are archived separately at https://data.ess-dive.lbl.gov/datasets/doi:10.15485/2428003 (Laan et al., 2024) and https://data.ess-dive.lbl.gov/datasets/doi:10.15485/1923689 (Forbes et al.,2023). A preliminary version of this data package was published in February 2026 at the time of manuscript submission. It was updated in June 2026, at the time of manuscript acceptance, to include the finalized data and additional metadata (readme, data dictionary, and file level metadata). For details on how to navigate data packages generated by this project, see https://data.ess-dive.lbl.gov/portals/PNNLRiverCorridorSFA/About. In addition to a readme, this data package also includes a file-level metadata (FLMD) file that describes each file and a data dictionary (DD) that describes all column/row headers and variable definitions. At the top level, the data package is organized into five main folders: (1) Data, (2)Figures, (3) Map, (4) GAM_Reulsts, and (5) src. The Data folder contains analysis-ready tabular files with oxygen consumption rates, DOM thermodynamic properties by site and treatment, site-level environmental variables, watershed-scale metrics, and other derived variables referenced in the manuscript. The Figures folder contains static image files associated with the main text and supplemental figures, while the Map folder includes spatial data and map-layer files used to create the sampling-location map. The GAM results folder contains the results for each of the general additive model (GAM).The src folder contains R scripts used to perform data processing, statistical analyses (including clustering, generalized additive models, and threshold analysis), and figure generation. This data package is associated with a GitHub repository found at https://github.com/WHONDRS-Hub/ECA_DOM_Thermodynamics.

Dissolved organic matter↗

Clustering, randomness and regularity in cloud fields. I - Theoretical considerations. II - Cumulus cloud fields

The current controversy existing in reference to the regularity vs. clustering in cloud fields is examined by means of analysis and simulation studies based upon nearest-neighbor cumulative distribution statistics. It is shown that the Poisson representation of random point processes is superior to pseudorandom-number-generated models and that pseudorandom-number-generated models bias the observed nearest-neighbor statistics towards regularity. Interpretation of this nearest-neighbor statistics is discussed for many cases of superpositions of clustering, randomness, and regularity. A detailed analysis is carried out of cumulus cloud field spatial distributions based upon Landsat, AVHRR, and Skylab data, showing that, when both large and small clouds are included in the cloud field distributions, the cloud field always has a strong clustering signal.

Weger, R. C.↗

BGC Atlas: a web resource for exploring the global chemical diversity encoded in bacterial genomes

Secondary metabolites are compounds not essential for an organism’s development, but provide significant ecological and physiological benefits. These compounds have applications in medicine, biotechnology and agriculture. Their production is encoded in biosynthetic gene clusters (BGCs), groups of genes collectively directing their biosynthesis. The advent of metagenomics has allowed researchers to study BGCs directly from environmental samples, identifying numerous previously unknown BGCs encoding unprecedented chemistry. Here, we present the BGC Atlas (https://bgc-atlas.cs.uni-tuebingen.de), a web resource that facilitates the exploration and analysis of BGC diversity in metagenomes. The BGC Atlas identifies and clusters BGCs from publicly available datasets, offering a centralized database and a web interface for metadata-aware exploration of BGCs and gene cluster families (GCFs). We analyzed over 35 000 datasets from MGnify, identifying nearly 1.8 million BGCs, which were clustered into GCFs. The analysis showed that ribosomally synthesized and post-translationally modified peptides are the most abundant compound class, with most GCFs exhibiting high environmental specificity. We believe that our tool will enable researchers to easily explore and analyze the BGC diversity in environmental samples, significantly enhancing our understanding of bacterial secondary metabolites, and promote the identification of ecological and evolutionary factors shaping the biosynthetic potential of microbial communities.

59 BASIC BIOLOGICAL SCIENCES↗