Search NASA⌕ Search

SEARCH · Search NASA

Results for “Classification”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 307 records · Page 17

An improved dataset for predicting mammal infecting viruses from genetic sequence information

There have been several attempts to develop machine learning (ML) models to identify human infecting viruses from their genomic sequences, with varying degrees of success. Direct comparison between models is problematic, because these models are typically trained and evaluated on different datasets with alternative data splitting schemes, features, and model performance metrics. In this paper we present a standardized dataset of mammal infecting and non-infecting viral pathogens, refined from the previous work of Mollentze et al. to include the latest literature evidence, roughly doubling the number of curated host-virus records available to the community, and new host target labels, primate and mammal. The new host labels were included for several reasons, including previous reports that classification performance is better at broader taxonomic ranks and the idea that there may be more data for primate infection that might serve as a suitable proxy for zoonotic potential and avoidance of false positives for human infection due to absence of evidence. On this dataset, we report the performance of eight machine learning models for predicting mammal-infecting viruses from their genomic sequences. We find that randomly assigning cases in our improved dataset to training/testing sets, when compared to the original assignments into training/testing in Mollentze et al., increases the overall average ROC AUC of prediction of human infection from 0.663 ± 0.070 to 0.784 ± 0.013, consistent with the reduction in phylogenetic distance between train and test sets (relative entropy change from 3.00 to 0.08). The broadest host category of mammal infection can be predicted most reliably at 0.850 ± 0.020. We share our improved dataset and code to enable standardized comparisons of machine learning methods to predict human host infections. Overall, we have presented preliminary evidence that classification of virus host infection is more tractable at higher taxonomic ranks, that unsurprisingly reducing the phylogenetic distance between training and test sets can improve predictive performance, that peptide kmer features appear to be harmful to out of sample model performance, and we are left with the question of whether models for virus host prediction can reasonably be expected to perform well in out of sample scenarios given the likelihood that viruses do not share a common ancestor. Consistent with this concern, when the data is resampled such that there is no overlap between viral families in training and test sets (relative entropy > 24), models perform no better than random chance at prediction of human infection regardless of whether kmers are included (ROC AUC 0.50 ± 0.08) or not (ROC AUC 0.50 ± 0.04).

59 BASIC BIOLOGICAL SCIENCES↗

Evaluation of WSR-88D Level III and MRMS Rainfall Estimates Against Rain Gauge Observations at the Savannah River Site

Rainfall data for the Savannah River Site (SRS) has been historically measured by rain gauges. These instruments serve as ground truth for most climatological and weather applications; however, gauge measurements are prone to errors or biases under certain weather conditions. Rainfall estimates from radar reflectivity values have been developed and improved over the years and serve as an alternative or supplement for gauge measurements. This study compares measurements from tipping bucket rain gauges with rainfall estimates from the NOAA NWS WSR-88D Level III hourly rainfall product and the Multi-Radar Multi-Sensor (MRMS) gauge-corrected precipitation estimate. Results show good agreement between radar derived amounts and ground measurements, with MRMS values showing correlation coefficients of 0.60-0.95 and RMSE values of less than 1.0 cm (0.4 in). The WSR-88D Level III estimates result in correlation coefficients of 0.46-0.86 and RMSE values of less than 1.3 cm (0.5 in). Few outliers are observed for each data pair and are evaluated against precipitation classification products (WSR-88D Hybrid Hydrometeor Classification product and MRMS Precipitation Flag product). The rain gauges used in this study are not part of the Hydrometeorological Automated Data System (HADS) network used to correct the MRMS estimates and therefore, results of this work provide an independent validation of the MRMS gauge correction scheme.

54 ENVIRONMENTAL SCIENCES↗

Hydrologic Model Data for the East Fork Poplar Creek Watershed Simulated with the Advanced Terrestrial Simulator (ATS): Streamflow and Network Expansion–Contraction Dynamics

This dataset supports hydrologic modeling and stream network expansion–contraction analysis for the East Fork Poplar Creek (EFPC) Watershed in Tennessee. It includes a Jupyter notebook for model setup, model configuration files, simulation outputs, and derived products used to evaluate model performance and investigate stream dynamics under varying hydrologic conditions. The dataset was generated using the Watershed Workflow Python package and the Advanced Terrestrial Simulator (ATS), enabling integrated surface–subsurface hydrologic simulations using a stream-aligned mesh. Outputs include high-resolution time series of streamflow, active network length, water table depth, and related hydrologic variables. Also included are spatially explicit stream persistency indices and classifications of reaches as perennial or non-perennial. These data facilitate reproducibility and support further research on stream intermittency and variability in network extent.The model data archive is organized in following directories:1) model_setup_inputsContains the Watershed Workflow Jupyter notebooks (accessed through any open source code editor), selected input datasets, and resulting ATS input files, including XML files (access through any open source code editor), computational mesh (.exo files can be viewed using Paraview), and meteorological forcing files (.h5 files can be accessed through h5py python package and HDFView open source software). 2) model_outputsIncludes ATS simulation outputs relevant to this study. Time series of spatially integrated or averaged variables (e.g., streamflow, water table depth) are provided as CSV files. Select spatial fields (e.g., ponded depth and water table depth) are saved as pickled Python objects to reduce file size, and can be accessed through pickle package in Python. Key geometry objects from Watershed Workflow—such as the surface mesh and river tree—are also included to support analysis of streamflow persistency and expansion–contraction dynamics. These files can also be accessed through Watershed Workflow Python package.3) model_evaluationProvides observed streamflow time series and field survey-based flow regime classifications used to evaluate model performance. Jupyter notebooks for processing ATS outputs and comparing model predictions with observations to build confidence in the model prior to scientific analysis are also included.4) Q_L_relationshipsContains workflows for generating time series of discharge, active network length, and related hydrologic variables used in the stream network expansion–contraction analysis. Includes routines for delineating baseflow-dominated periods. For each catchment, notebooks and processed data (as pickled DataFrames accessed through Pandas Python package) are provided. 5) figure_scriptsProvides the Jupyter notebooks used to generate the figures presented in the paper.

54 ENVIRONMENTAL SCIENCES↗

Exploring Data Set Bias and Decision Support with Predictive Uncertainty Through Bayesian Approximations and Convolutional Neural Networks

Individual seismic catalogs can contain multiscale observations from fault level to global scales and associated waveforms from discrete events reflect crustal structure across many different scales and locations. Seismic network aperture, geographic location, and observation distance may not provide informative guidance or intuition on how different catalogs will behave across models trained under different conditions. We rely on uncertainty to provide guardrails for when to trust model decisions, but understanding when our uncertainty is trustworthy is an open challenge. Here, in this work, we explore Bayesian approximation methods for assigning predictive uncertainty in seismic event classification problems. We find that computationally expensive Bayesian approximations do not outperform simple ensemble methods. We also find that when exploiting multiple seismic event catalogs, joint training with data from all the catalogs combined with Bayesian approximations and supervised training for classification can obscure bias and result in less robust uncertainty while also not providing substantial performance benefits compared to training individual models for each catalog.

58 GEOSCIENCES↗

Neutron-antineutron oscillation sensitivity study at DUNE

The Deep Underground Neutrino Experiment (DUNE) aims to measure neutrino oscillations as well as search for beyond the standard model physics such as baryon number violating (BNV) processes. DUNE will use a 70 kt Liquid Argon Time Projection Chamber (LArTPC) located more than 1 km underground. A promising BNV process is neutron-antineutron oscillation ($n \rightarrow \bar{n}$) which, if discovered, would offer unique insight into the baryon asymmetry of the universe. We are developing a classification algorithm that separates $n \rightarrow \bar{n}$ events from major background atmospheric neutrino interactions using DUNE far detector simulations. We will perform the classification of signals and backgrounds by analyzing key features such as the multiplicity, isotropy, and kinematics of the reconstructed events. In the future, this algorithm can be used to obtain the sensitivity of the DUNE detectors to the neutron-antineutron oscillation lifetime.

Wheeler, Justin↗

Woody Feedstock 2022 State of Technology Report

The U.S. Department of Energy promotes production of advanced liquid transportation fuels from lignocellulosic biomass by funding fundamental and applied research that advances the state of technology (SOT). As part of its involvement in this mission, Idaho National Laboratory completes an annual SOT report for n th -plant and 1 st -plant woody biomass feedstock logistics. The purpose of the SOT is to provide the status of feedstock supply system technology development for woody biomass to biofuels relative to technical targets and cost goals from specific design cases, based on data and experimental results. Conventional feedstock supply systems need to be modified to meet the demands of conversion pathways, specifically to have the ability to adjust the quality of the raw biomass materials. Advanced systems incorporate innovative methods of material handling, preprocessing and supply chain configuration. In advanced designs, variability of the raw biomass can be reduced to produce feedstocks of a uniform format, moving toward biomass commoditization. Against this backdrop, the 2022 Woody SOT for low-ash woody feedstocks utilizes feedstock fractionation by incorporating technologies that can separate the biomass into its anatomical fractions (wood, bark, needle, and extrinsic ash) to reduce impurities and attempt to maximize the retention of usable fractions that satisfy downstream quality considerations. By using a series of air classification steps, this strategy can reduce the extrinsic ash in forest residues, separate out a majority of the incoming needles (which can be supplied to alternate markets), and maximize the retention of whitewood in the usable fraction. The fractionated forest residues are then mixed with clean-pine chips in a 50-50 blend to prepare the feedstock for the desired conversion pathway. The n th -plant analysis estimated the delivered cost for the feedstock at $\$$69.23/dry ton (2016$\$$) which represents a $\$$6.64/dry ton decrease compared to the cost estimate of the 2021 Woody SOT supply system for low-ash woody feedstocks. The quality requirements in the 2022 Woody SOT were identical to those of the 2021 Woody SOT at = 1.00 wt % ash and = 50.51 wt% carbon. The cost savings derive primarily from reductions in dry matter losses during air classification. The GHG emissions for the n th -plant analysis were estimated at 178.39 kg CO2e/dry ton compared to 178.71 kg CO 2 e/dry ton in the 2021 Woody SOT, a decrease of 0.32 kg CO2e/dry ton. The small change stems from an increase in emissions attributed to preprocessing and slightly larger savings in emissions from transportation. In the 1 st -plant analysis of the 2022 Woody SOT system, the average throughput was estimated to be approximately 2,128 dry tons/day or 96.51% of the name plate capacity. During the simulation the daily throughput ranged from 1,090 dry tons/day to 2,200 dry tons/day, or 49.43% to 99.75% of the daily nameplate capacity. After the year of operation 722,403 tons of processed feedstock were produced in total without regard to quality considerations (99.64% of the annual nameplate capacity). The variability in throughput was primarily caused by equipment failures in the system. Regular failures, downtime caused by routine maintenance per manufacturer guidelines, contributed to a majority 62.50% of failures and 62.60% of downtime. Failures due to wear were the other cause of disruption within the system, impacting the rotary shear and orbital screen and accounting for 37.50% of the failures and 37.40% of the total downtime. Ultimately the system was on stream for 87.84% during the simulation period, which is only 2.16 percentage points below the nth-plant assumption for on-stream time. The production cost of the system averaged $\$$71.66/dry ton. The costs ranged from a minimum of $\$$71.23/dry ton to a maximum of $\$$2,115.30/dry ton. When dry matter losses (disposed low-quality fractions as well as other losses such as in grinders) were considered the costs increased to an average of $\$$75.11/dry ton with a minimum of $\$$74.69/dry ton and a maximum of $\$$2,136.86/dry ton...

09 BIOMASS FUELS↗

Radioisotope Identification with List-Mode Gamma Ray Data: A rigorous assessment on the value of temporal information applied to radioisotope identification.

This work explores the potential of utilizing temporal data from gamma-ray detectors, known as list-mode data, to enhance radioisotope identification. Traditional identification methods, which rely on full gamma-ray spectrum analysis, often require long dwell times and struggle with “confuser” sources, or spectra with similarly spaced spectral peaks. We hypothesize that by leveraging the probabilistic nature of nuclear decay and the time-encoded information from decay sequences and interactions with surrounding materials, we can improve classification accuracy over static spectral analysis. This research rigorously examines the temporal content of list-mode data through exploratory data analysis via correlation discovery and information theory. We further propose a basic classification model that can utilize spectral or temporal data (or both) to determine if the incorporation of temporal information can improve radioisotope identification. The findings suggest that the temporal information present in list-mode gamma-ray data has merit and should be further investigated.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Passband Signal Detection at the Edge

Algorithms for radio frequency (RF) spectrum awareness need to be compatible with edge hardware to be practical for many applications. We developed a signal detection and classification model for the ZCU111 RF System-on-a-Chip (RFSoC) that operates on the fast Fourier transform of passband RF data. The system can detect and classify multiple signals of interest and display the predictions in real-time. The model consists of a modified ConvNeXt backbone and YOLOv3 head to operate on the Deep Learning Processing Unit on the RFSoC. We gathered datasets for training and testing by using a software defined radio to transmit example signals of Wi-Fi 802.11 b/g, Wi-Fi 802.11 n, FM Radio, LTE and LTE-M. By leveraging multiple inputs on the RFSoC frontend, the datasets span up to 4 GHz of bandwidth. The models showed high performance in classification accuracy, center frequency error, bandwidth error, and detection accuracy for both single and multi-signal datasets.

42 ENGINEERING↗

Deploying Adversarial Attacks in Super-Resolution Models

Reliable super-resolution methods are crucial for applications like remote sensing, grid resilience and disaster impact analysis, and standoff biometrics. These methods infuse additional high-frequency information into reconstructions, allowing for better contextualization and image intelligence. However, super-resolution models can also introduce hallucinations or other unseen vulnerabilities that could be exploited by an adversary. This is further compounded by the prominence of deep learning in these models, as models are often blindly applied on out-of-distribution images. In this work, we implement adversarial attacks in common open-source super-resolution models and examine their impact on reconstructions and downstream classification tasks. We find that an adversarially trained super-resolution model can produce high-quality reconstructions that degrade downstream classifications. Moreover, these attacks do not require access to low-resolution imagery or class labels at inference time. These results demonstrate the vulnerability of super-resolution methods to malicious actors and motivates the development of a detector for super-resolution adversarial attacks. Further exploration of adversarial attacks in this domain is required to ensure trustworthiness and robustness of super-resolution models for national security applications.

45 MILITARY TECHNOLOGY, WEAPONRY, AND NATIONAL DEF↗

Assessing Dynamic Time Warping Techniques for Discriminating Seismic Sources at Local and Regional Distances

Effective monitoring of seismic explosions and hazard assessment relies heavily on the accurate discrimination of underground seismic sources. This study investigates the application of novel nonlinear alignment techniques, specifically Dynamic Time Warping (DTW), for event-type discrimination at regional and local distances. Building on prior research that used DTW and Elastic Shape Analysis (ESA) in discrimination at regional distances, we evaluate the performance of recently developed variants of DTW, including a method that employs Pearson cross-correlation as a measure of warping distance and a time distortion coefficient that quantifies the type and degree of time distortion between signals. By analyzing observational datasets that include different source types, we assess the performance of these approaches for realistic monitoring scenarios. Specifically, we consider a dataset recorded at regional distances in the Korean Peninsula and a local-distance subset from the Unconstrained Utah Event Bulletin catalog to evaluate DTW-based discrimination across multiple distance scales. Additionally, we introduce the maximum cross-correlations of warped waveforms as a similarity metric for event classification. Through hierarchical cluster analysis and dendrogram interpretation, we present our findings, highlighting the strengths and limitations of these techniques in seismic event classification.

45 MILITARY TECHNOLOGY, WEAPONRY, AND NATIONAL DEF↗

Predictions for the Detectability of Milky Way Satellite Galaxies and Outer-Halo Star Clusters with the Vera C. Rubin Observatory

We predict the sensitivity of the Vera C. Rubin Observatory Legacy Survey of Space and Time (LSST) to faint, resolved Milky Way satellite galaxies and outer-halo star clusters. We characterize the expected sensitivity using simulated LSST data from the LSST Dark Energy Science Collaboration (DESC) Data Challenge 2 (DC2) accessed and analyzed with the Rubin Science Platform as part of the Rubin Early Science Program. We simulate resolved stellar populations of Milky Way satellite galaxies and outer-halo star clusters over a wide range of sizes, luminosities, and heliocentric distances, which are broadly consistent with expectations for the Milky Way satellite system. We inject simulated stars into the DC2 catalog with realistic photometric uncertainties and star/galaxy separation derived from the DC2 data itself. We assess the probability that each simulated system would be detected by LSST using a conventional isochrone matched-filter technique. We find that assuming perfect star/galaxy separation enables the detection of resolved stellar systems with $M_V$ = 0 mag and $r_{1/2}$ = 10 pc with >50% efficiency out to a heliocentric distance of ~250 kpc. Similar detection efficiency is possible with a simple star/galaxy separation criterion based on measured quantities, although the false positive rate is higher due to leakage of background galaxies into the stellar sample. When assuming perfect star/galaxy classification and a model for the galaxy-halo connection fit to current data, we predict that 89 +/- 20 Milky Way satellite galaxies will be detectable with a simple matched-filter algorithm applied to the LSST wide-fast-deep data set. Different assumptions about the performance of star/galaxy classification efficiency can decrease this estimate by ~7%-25%, which emphasizes the importance of high-quality star/galaxy separation for studies of the Milky Way satellite population with LSST.

79 ASTRONOMY AND ASTROPHYSICS↗

Embedded Sensing in Additive Manufacturing Metal and Polymer Parts: A Comparative Study of Integration Techniques and Structural Health Monitoring Performance

This study presents a comparative evaluation of post-process sensor integration in additively manufactured (AM) metal and the in-situ process for polymer structures for structural health monitoring (SHM), with an emphasis on embedded sensors. Geometrically identical specimens were fabricated using copper via metal fused filament fabrication (FFF) and PLA via polymer FFF, with piezoelectric transducers (PZTs) inserted into internal cavities to assess the influence of material and placement on sensing fidelity. Mechanical testing under compressive and point loads generated signals that were transformed into time–frequency spectrograms using a Short-Time Fourier Transform (STFT) framework. An engineered RGB representation was developed, combining global amplitude scaling with an amplitude-envelope encoding to enhance contrast and highlight subtle wave features. These spectrograms served as inputs to convolutional neural networks (CNNs) for classification of load conditions and detection of damage-related features. Results showed reliable recognition in both copper and PLA specimens, with CNN classification accuracies exceeding 95%. Embedded PZTs were especially effective in PLA, where signal damping and environmental sensitivity often hinder surface-mounted sensors. This work demonstrates the advantages of embedded sensing in AM structures, particularly when paired with spectrogram-based feature engineering and CNN modeling, advancing real-time SHM for aerospace, energy, and defense applications.

additive manufacturing↗

Surface-Enhanced Raman Spectroscopy Combined with Multivariate Analysis for Fingerprinting Clinically Similar Fibromyalgia and Long COVID Syndromes

Fibromyalgia (FM) is a chronic central sensitivity syndrome characterized by augmented pain processing at diffuse body sites and presents as a multimorbid clinical condition. Long COVID (LC) is a heterogenous clinical syndrome that affects 10–20% of individuals following COVID-19 infection. FM and LC share similarities with regard to the pain and other clinical symptoms experienced, thereby posing a challenge for accurate diagnosis. This research explores the feasibility of using surface-enhanced Raman spectroscopy (SERS) combined with soft independent modelling of class analogies (SIMCAs) to develop classification models differentiating LC and FM. Venous blood samples were collected using two supports, dried bloodspot cards (DBS, n = 48 FM and n = 46 LC) and volumetric absorptive micro-sampling tips (VAMS, n = 39 FM and n = 39 LC). A semi-permeable membrane (10 kDa) was used to extract low molecular fraction (LMF) from the blood samples, and Raman spectra were acquired using SERS with gold nanoparticles (AuNPs). Soft independent modelling of class analogy (SIMCA) models developed with spectral data of blood samples collected in VAMS tips showed superior performance with a validation performance of 100% accuracy, sensitivity, and specificity, achieving an excellent classification accuracy of 0.86 area under the curve (AUC). Amide groups, aromatic and acidic amino acids were responsible for the discrimination patterns among FM and LC syndromes, emphasizing the findings from our previous studies. Overall, our results demonstrate the ability of AuNP SERS to identify unique metabolites that can be potentially used as spectral biomarkers to differentiate FM and LC.

60 APPLIED LIFE SCIENCES↗

The Zwicky Transient Facility Bright Transient Survey. III. BTSbot: Automated Identification and Follow-up of Bright Transients with Deep Learning

Abstract The Bright Transient Survey (BTS) aims to obtain a classification spectrum for all bright ( m peak ≤ 18.5 mag) extragalactic transients found in the Zwicky Transient Facility (ZTF) public survey. BTS critically relies on visual inspection (“scanning”) to select targets for spectroscopic follow-up, which, while effective, has required a significant time investment over the past ∼5 yr of ZTF operations. We present BTSbot , a multimodal convolutional neural network, which provides a bright transient score to individual ZTF detections using their image data and 25 extracted features. BTSbot is able to eliminate the need for daily human scanning by automatically identifying and requesting spectroscopic follow-up observations of new bright transient candidates. BTSbot recovers all bright transients in our test split and performs on par with scanners in terms of identification speed (on average, ∼1 hr quicker than scanners). We also find that BTSbot is not significantly impacted by any data shift by comparing performance across a concealed test split and a sample of very recent BTS candidates. BTSbot has been integrated into Fritz and Kowalski , ZTF’s first-party marshal and alert broker, and now sends automatic spectroscopic follow-up requests for the new transients it identifies. Between 2023 December and 2024 May, BTSbot selected 609 sources in real time, 96% of which were real extragalactic transients. With BTSbot and other automation tools, the BTS workflow has produced the first fully automatic end-to-end discovery and classification of a transient, representing a significant reduction in the human time needed to scan.

Rehemtulla, Nabeel (ORCID:0000000256832389)↗

Distinguishing Orbiting and Infalling Dark Matter Particles with Machine Learning

Dark matter halos are typically defined as spheres that enclose some overdensity, but these sharp, somewhat arbitrary boundaries introduce nonphysical artifacts such as backsplash halos, pseudo-volution, and an incomplete accounting of halo mass. A more physically motivated alternative is to define halos as the collection of particles that are physically orbiting within their potential well. However, existing methods to classify particles as orbiting or infalling suffer from trade-offs between accuracy, computational cost, and generalizability across cosmologies. We present an efficient, yet accurate, supervised machine learning approach using decision trees. The classification is based on only the particle radii and velocities at two epochs. Compared to detailed analysis of particle trajectories, we find that our model matches the classification of 97% of particles. Consequently, we are able to quickly and accurately reproduce the density profiles of the orbiting and infalling components out to many virial radii. We demonstrate that our model generalizes to a significantly different cosmology that lies outside the training data set. We make publicly available both our final model and the code to train similar models.

79 ASTRONOMY AND ASTROPHYSICS↗

The MOST Hosts Survey: Spectroscopic Observation of the Host Galaxies of ∼40,000 Transients Using DESI

We present the Multi-Object Spectroscopy of Transient (MOST) Hosts survey. The survey is planned to run throughout the 5 yr of operation of the Dark Energy Spectroscopic Instrument (DESI) and will generate a spectroscopic catalog of the hosts of most transients observed to date, in particular all the supernovae observed by most public, untargeted, wide-field, optical surveys (Palomar Transient Factory, PTF/intermediate PTF, Sloan Digital Sky Survey II, Zwicky Transient Facility, DECAT, DESIRT). Science cases for the MOST Hosts survey include Type Ia supernova cosmology, fundamental plane and peculiar velocity measurements, and the understanding of the correlations between transients and their host-galaxy properties. Here we present the first release of the MOST Hosts survey: 21,931 hosts of 20,235 transients. These numbers represent 36% of the final MOST Hosts sample, consisting of 60,212 potential host galaxies of 38,603 transients (a transient can be assigned multiple potential hosts). Of all the transients in the MOST Hosts list, only 26.7% have existing classifications, and so the survey will provide redshifts (and luminosities) for nearly 30,000 transients. A preliminary Hubble diagram and a transient luminosity–duration diagram are shown as examples of future potential uses of the MOST Hosts survey. The survey will also provide a training sample of spectroscopically observed transients for classifiers relying only on photometry, as we enter an era when most newly observed transients will lack spectroscopic classification. The MOST Hosts DESI survey data will be released on a rolling cadence and updated to match the DESI releases.

79 ASTRONOMY AND ASTROPHYSICS↗

Dark Energy Survey Year 6 Results: Photometric Dataset for Cosmology

We describe the photometric dataset assembled from the full 6 yr of observations by the Dark Energy Survey (DES) in support of static-sky cosmology analyses. DES Y6 Gold is a curated dataset derived from DES Data Release 2 (DR2) that incorporates improved measurement, photometric calibration, object classification and value-added information. Y6 Gold comprises nearly 5000 deg$^{2}$ of grizY imaging in the south Galactic cap and includes 669 million objects with a depth of i$_{AB}$ ∼ 23.4 mag at a signal-to-noise ratio ∼ 10 for extended objects and a top-of-the-atmosphere photometric uniformity <2 mmag. Y6 Gold augments DES DR2 with simultaneous fits to multiepoch photometry for more robust galaxy shapes, colors, and photometric redshift estimates. Y6 Gold features improved morphological star–galaxy classification with an efficiency of 98.6% and a contamination of 0.8% for galaxies with 17.5 < i$_{AB}$ < 22.5. Additionally, it includes per-object quality information, and accompanying maps of the footprint coverage, masked regions, imaging depth, survey conditions, and astrophysical foregrounds that are used for cosmology analyses. After quality selections, benchmark samples contain 448 million galaxies and 120 million stars. This publication is complemented by data access and documentation.

79 ASTRONOMY AND ASTROPHYSICS↗

Applying a Multisector Scenario Framework to Evaluate Past and Future Public Surface Water Supply Infrastructure Strategies in Texas

Datasets supporting the index model and scenario analysis used in evaluating surface water supply strategies across different water system types in Texas. These data underpin the scenario development and application of five key indicators: Water Availability Index (WAI), Water Quality Index (WQI), Energy Requirement Index (ERI), Water Treatment Cost (WTC), and Water Infrastructure Cost (WIC). The datasets are organized by system type—stream reaches (flowlines), waterbodies, and reservoirs—and include both raw and standardized index values. The integrated datasets also provide scenario classifications (original and adjusted) based on infrastructure and planning priorities, enabling comparison across Shared Socioeconomic Pathways (SSPs). Additional strategy-level data are included to support evaluation of state-level new reservoir projects in relation to cost and availability tradeoffs. Please refer to the README file provided in Files for more details. Descriptions of the datasets are provided below. Dataset(s) Descriptions Folder: Index_model_database.zip Subfolder: Stream_reach.zip Fl_wf.csv, Fl_wq.csv, Fl_er.csv, Fl_wf_wtcUV.csv, Fl_wf_wtcnoUV.csv, Fl_allfac_wic1.csv, Fl_allfac_wic2.csvDatasets for computing WAI, WQI, ERI, WTC, and WIC for surface water systems classified as stream reaches (flowlines). Subfolder: Waterbody.zip Wb_wf.csv, Wb_wq.csv, Wb_er.csv, Wb_wf_wtcUV.csv, Wb_wf_wtcnoUV.csv, Wb_allfac_wic1.csv, Wb_allfac_wic2.csvEquivalent index model datasets for waterbodies, reflecting hydrologic and infrastructure attributes specific to impounded natural systems. Subfolder: Reservoir.zip Rs_wf.csv, Rs_wq.csv, Rs_er.csv, Rs_wf_wtcUV.csv, Rs_wf_wtcnoUV.csv, Rs_allfac_wic1.csv, Rs_allfac_wic2.csvIndex model datasets specific to regulated reservoir systems, incorporating both resource indicators and cost parameters. Folder: Integrated data.zip combined_merged_data.csv, combined_merged_data_scenario.csvDatasets integrating index model indicators (both raw and scaled) with scenario classifications, including adjustments reflecting SSP-aligned transitions and planning shifts. Folder: Additional data.zip wai_supplystrat_wic_merged.csvCurated dataset capturing proposed major reservoir-based municipal water supply strategies in Texas. Integrates site-level planning data with estimated capital infrastructure costs and water availability scores for comparative assessment.

geospatial↗