Search NASA⌕ Search

SEARCH · Search NASA

Results for “Streaming data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 307 records · Page 17

Pennsylvania Department of Environmental Protection (PA DEP) 26r Detailed Produced Water Compositions (version 1.0)

A database of geochemical compositions of aqueous species in produced water reported to the PA DEP. Samples were collected between mid-2012 to early-2020. Data from publicly-available PA DEP 26r reports were scraped from pdf files and cumulated into tabular spreadsheet format for >1000 produced water streams from Marcellus wells in Pennsylvania. In addition to providing the original values, the NETL NEWTS team has reformatted the dataset to allow sample streams to be easily copied into OLI Studio and Geochemist WorkBench (GWB) software for modeling the geochemistry and the recovery of critical minerals, such as lithium, from these produced water streams. In addition, a version of the dataset has been included with predictions for some missing values in the original dataset using machine learning techniques within CoDaRT software, a public ML software developed by the Nation Energy Technology Laboratory. We have made the Input into CoDaRT and one example output from CoDaRT available in this dataset.

Aqueous Chemistry↗

Pennsylvania Department of Environmental Protection (PA DEP) 26r Detailed Produced Water Compositions (version 2.0)

A database of geochemical compositions of aqueous species in produced water reported to the PA Department of Environmental Protection (PA DEP). Samples were collected between late-2010 to late-2024. Data from publicly available PA DEP 26r reports were scraped from pdf files and cumulated into tabular spreadsheet format for >3,000 produced water streams from Marcellus Shale wells in Pennsylvania. In addition to providing the original values, the NETL NEWTS team has reformatted the dataset to allow sample streams to be easily copied into OLI Studio and Geochemist WorkBench (GWB) software for modeling the geochemistry and the recovery of critical minerals, such as lithium, from these produced water streams. ***This dataset is an updated version of the PA DEP 26r Detailed Produced Water Compositions (version 1.0) dataset, providing expanded spatial and temporal coverage.***

Aqueous Chemistry↗

Network of Modified Portable Optical Particle Spectrometers Instrument Handbook

The U.S. Department of Energy Atmospheric Radiation Measurement (ARM) user facility’s portable optical particle spectrometer (POPS) network (Figure 1) provides an aerosol size distribution spectrum from 150 nm to 5 μm in 16 logarithmically spaced bins. The data emanating from this network of up to four matching POPS in weatherproof enclosures comprises one-minute averages of the aerosol size and number spectrum and the co-located temperature and relative humidity. The total particle concentration and sample flow rate are also provided. The sample stream is not dried before its measurement. Two size spectra are provided: the default size distribution corresponds to the calibration using polystyrene latex spheres (PSL), while the second size distribution corresponds to the calibration using size-selected dry ammonium sulfate aerosol. The principal scientific application of the data provided by the ARM POPS is to determine the aerosol size distribution across a region in three disparate locations – the size distribution is broadly relevant to studies of pollution and of conditions in the boundary layer. A high particle concentration can be an indicator of unhealthy levels of pollution; the dispersion of particles in the atmosphere is a predictor of long-range visibility; and particles serve as collectors of condensing gases including water vapor, and thus are a factor in the formation and properties of fogs and clouds, and in the initiation of precipitation.

54 ENVIRONMENTAL SCIENCES↗

Field and Model Data Associated with the Manuscript “Drivers of Streamflow Intermittency in Humid Regions: 1. Evaluating Above- and Below-ground Controls of Flow Persistence in a Forested Catchment”

This package contains field data, modeling files, and scripts supporting the investigation of the drivers of streamflow intermittency in a forested catchment. It includes the field data collected from electrical resistivity tomography (ERT) surveys, ground penetrating radar (GPR), continuous self-potential (SP) monitoring, electromagnetic (EM) imaging, groundwater and stilling well. In addition, it contains the data and results of the coupled water- and electrical-flow model developed using the COMSOL Multiphysics and Advanced Terrestrial Simulator (ATS), as well as software files and Jupyter notebooks used to process the data and generate figures in the manuscript submitted for peer review. The data archive is organized in the following directories: 1) Climate Includes hourly precipitation and daily evapotranspiration time series (2024 – 2025) provided as CSV files, alongside a text file detailing dataset units. 2) Coupled_model Contains two subfolders: Synthetic and Field_Application subfolder. Synthetic subfolder contains the ATS XML input script (can be opened using any code editor) for the four synthetic hydrological cases tested (Connected and gaining, Connected and losing, Disconnected and losing, and dry stream). It also includes other experimental cases to test the influence of precipitation and concentration gradient. For each synthetic case, the flow model simulation is executed using the ATS XML scripts and the included Python script (generate_data_set.py) to convert ATS output to COMSOL-ready input. COMSOL Multiphysics template (.mph can be opened with the commercial software COMSOL and requires a license) is executed using the ATS output data to simulate the potential field. It also includes the Synthetic_model_plot.ipynb (can be opened using any code editor) to visualize the SP result and generate manuscript figures. The data subfolder contains mesh files to run both the ATS (.exo and .stl files can be viewed using Paraview; .h5 files can be opened using HDFView software and h5py Python package) and COMSOL models. Field_Application subfolder contains two subfolders: ES_MDA_inversion and Final_Model. ES_MDA_inversion contains the Python script (.py can be opened using any code editor) and SP observation data used to run the Ensemble Smoother with Multiple Data Assimilation (ES-MDA) inversion sequence to get the optimal model parameters. The Final_model subfolder contains the ATS XML input scripts, data files, output data for the two SP sites. The same workflow steps outlined for the Synthetic subfolder apply here. It also contains the Jupyter notebook (Plot_final_calib.ipynb) to visualize the results of the modeled SP, stream-groundwater exchange and moisture content. 3) Discharge Includes the electrical conductivity (EC) time series (provided as CSV files) from salt slug injections. It also includes the Jupyter notebook (Discharge_process.ipynyb) used to estimate discharge. All discharge measurements collated into rating_curve_processed.csv 4) EM Contains the CSV file of the EM data from the DUALEM-42, including spatial coordinates (x, y, z), apparent conductivity, and in-phase measurements at 2 m coil separations for horizontal coplanar (HCP) and perpendicular (PRP) geometries. 5) ERT Contains raw resistivity data (provided as CSV files), spatial location of each of the electrodes (provided as CSV files), and files used for the resistivity inversion (.resipy can be opened with the open-source ResIPy software). 6) GPR Includes GPR field datasets collected at 100 MHz and 250 MHz antenna frequencies, along with the processing/interpretation project file (GPR_process.gpz can be viewed using EKKO_Project 6, a commercial software by Sensors & Software that requires a license). 7) Slug_test Includes the slug test data at all the groundwater wells provided as CSV files, as well as the Jupyter notebook (Slug_test.ipynb) for calculating hydraulic conductivity. 8) SP Contains the SP data collected in field at the two SP sites (one in the perennial reach and the other in the intermittent reach), provided as DAT files. 9) Well_data Contains two subfolders: 1) Raw, which provides unprocessed pressure, electrical conductivity and temperature timeseries downloaded from the loggers in all the groundwater and stilling wells, and 2) Processed, which contains sorted, QA/QC timeseries data for each well. The data archive also contains data_process.ipynb, a Jupyter notebook used for field data analysis and generating figures (plotting well, SP, climate, and discharge data, as well as calculating head gradient at sites with nested groundwater wells). It also includes DTW.ipynb, a Jupyter notebook containing the code for the dynamic time warping (DTW) with sliding window to evaluate SP signal synchronicity.

ATS↗

Assessing Heterogeneity of Surface Water Temperature Following Stream Restoration and a High-Intensity Fire from Thermal Imagery

Thermal heterogeneity of rivers is essential to support freshwater biodiversity. Salmon behaviorally thermoregulate by moving from patches of warm water to cold water. When implementing river restoration projects, it is essential to monitor changes in temperature and thermal heterogeneity through time to assess the impacts to a river’s thermal regime. Lightweight sensors that record both thermal infrared (TIR) and multispectral data carried via unoccupied aircraft systems (UASs) present an opportunity to monitor temperature variations at high spatial (<0.5 m) and temporal resolution, facilitating the detection of the small patches of varying temperatures salmon require. Here, we present methods to classify and filter visible wetted area, including a novel procedure to measure canopy cover, and extract and correct radiant surface water temperature to evaluate changes in the variability of stream temperature pre- and post-restoration followed by a high-intensity fire in a section of the river corridor of the South Fork McKenzie River, Oregon. We used a simple linear model to correct the TIR data by imaging a water bath where the temperature increased from 9.5 to 33.4 °C. The resulting model reduced the mean absolute error from 1.62 to 0.35 °C. We applied this correction to TIR-measured temperatures of wetted cells classified using NDWI imagery acquired in the field. We found warmer conditions (+2.6 °C) after restoration (p < 0.001) and median absolute deviation for pre-restoration (0.30) to be less than both that of post-restoration (0.85) and post-fire (0.79) orthomosaics. In addition, there was statistically significant evidence to support the hypothesis of shifts in temperature distributions pre- and post-restoration (KS test 2009 vs. 2019, p < 0.001, D = 0.99; KS test 2019 vs. 2021, p < 0.001, D = 0.10). Moreover, we used a Generalized Additive Model (GAM) that included spatial and environmental predictors (i.e., canopy cover calculated from multispectral NDVI and photogrammetrically derived digital elevation model) to model TIR temperature from a transect along the main river channel. This model explained 89% of the deviance, and the predictor variables showed statistical significance. Collectively, our study underscored the potential of a multispectral/TIR sensor to assess thermal heterogeneity in large and complex river systems.

Barker, Matthew I. (ORCID:0000000252864930)↗

Enhancing Streamflow Reanalysis Across the Conterminous US Leveraging Multiple Gridded Precipitation Data Sets

Streamflow observations, essential for various water resource applications, are often unavailable at critical locations in need. Although different models have been proposed to enhance streamflow predictability at ungauged locations, the challenge extends beyond model fidelity. Differences in meteorologic forcing data sets, precipitation in particular, can significantly affect the accuracy of hydrologic predictions. This challenge intensifies across regions characterized by diverse hydro-climatological and geographical conditions, such as in the conterminous US (CONUS) where a single precipitation product struggles to consistently replicate observed hydrographs, particularly peak flow dynamics. To enhance streamflow predictions, we utilize a VIC-RAPID hydrologic modeling framework driven by multiple commonly used meteorological forcing data sets, such as Daymet, PRISM, ST4, AORC, and their hybrids and create multiple sets of 40-year (1980–2019) hourly, daily, and monthly streamflow reanalysis, Dayflow Version 2, for 2.7 million river reaches across the CONUS. Most forcings lead to skillful streamflow performance, except for ST4 in the mountainous west, where severe radar blockage adversely affects the accuracy. The evaluation using over 6,000 hourly stream gauges shows that hourly AORC and ST4 lead to improved annual peak flow performance over Daymet—driven streamflow (Dayflow V1), particularly in smaller basins, highlighting the value of high temporal resolution forcings in hydrologic predictions. Compared with other benchmark data sets like National Water Model V3.0, AORC-driven VIC-RAPID exhibits improved regional streamflow performance, with comparable peak flow representation. We envision that multi-forcing streamflow reanalysis data can inform regions in need of forcing data enhancement, diagnose hydrologic model performance, and benefit diverse water resource applications.

54 ENVIRONMENTAL SCIENCES↗

A projection method for particle resampling

Particle discretizations of partial differential equations are advantageous for high-dimensional kinetic models in phase-space due to their better scalability than continuum approaches with respect to dimension. Complex processes collectively referred to as particle noise hamper long time simulations with particle methods. One approach to address this problem is particle mesh adaptivity, or remapping, known as particle resampling and remeshing. Here, this work introduces a resampling method that projects particles to and from a (finite element) function space. The method is simple, using standard sparse linear algebra and finite element techniques, and it preserves all moments up to the order of a polynomial represented exactly by the continuum function space. It is distinguished from most other mesh-based methods in that new particle positions and number are decoupled from the mesh, allowing particle and continuum meshes to be adapted relatively independently. While this work is developed with structured particle and continuum phase-space grids on 1X + 1V Vlasov-Poisson models of Landau damping and two-stream instability, the method is well-suited to unstructured grids. Stable long time dynamics are demonstrated up to time T = 500. Reproducibility artifacts and data are publicly available.

Kinetic methods↗

Time Series Foundation Models and Deep Learning Architectures for Earthquake Temporal and Spatial Nowcasting

Advancing the capabilities of earthquake nowcasting, the real-time forecasting of seismic activities, remains crucial for reducing casualties. This multifaceted challenge has recently gained attention within the deep learning domain, facilitated by the availability of extensive earthquake datasets. Despite significant advancements, the existing literature on earthquake nowcasting lacks comprehensive evaluations of pre-trained foundation models and modern deep learning architectures; each focuses on a different aspect of data, such as spatial relationships, temporal patterns, and multi-scale dependencies. This paper addresses the mentioned gap by analyzing different architectures and introducing two innovative approaches called Multi Foundation Quake and GNNCoder. We formulate earthquake nowcasting as a time series forecasting problem for the next 14 days within 0.1-degree spatial bins in Southern California. Earthquake time series are generated using the logarithm energy released by quakes, spanning 1986 to 2024. Our comprehensive evaluations demonstrate that our introduced models outperform other custom architectures by effectively capturing temporal-spatial relationships inherent in seismic data. The performance of existing foundation models varies significantly based on the pre-training datasets, emphasizing the need for careful dataset selection. However, we introduce a novel method, Multi Foundation Quake, that achieves the best overall performance by combining a bespoke pattern with Foundation model results handled as auxiliary streams.

97 MATHEMATICS AND COMPUTING↗

U.S. Hydropower Development Pipeline Data, 2026

The U.S. Hydropower Development Pipeline dataset provides a comprehensive, regularly updated view of proposed and potential hydropower projects across the United States. This resource compiles information from federal agencies and other public sources to track non-powered dams considered for electrification, proposed hydropower facilities at stream reaches with no existing dams, conduit exemptions, and emerging pumped storage hydropower proposals. The dataset includes project characteristics such as location, development status, technology type, ownership category, and other attributes that support analysis of future hydropower trends. It is designed to help researchers, planners, policymakers, and stakeholders assess national‑scale development patterns, understand the evolving hydropower landscape, and explore opportunities and challenges associated with new hydropower deployment. The dataset is updated annually to reflect changes in project status, new proposals entering the pipeline, and projects that are cancelled, completed, or otherwise removed from active consideration. Note: Capacity additions to existing hydropower plants are not included in this database due to reliance on a proprietary data source.

Johnson, Megan [ORNL] (ORCID:0000000290141741)↗

Forecasting Dark Matter Subhalo Constraints from Stellar Streams using Implicit Likelihood Inference

The evidence for dark matter (DM) remains compelling, although attempts to understand its particle nature remain inconclusive. One promising method to study DM is detecting DM subhalos through their gravitational interactions with stellar streams. In this study, we apply Neural Posterior Estimation (NPE) to constrain subhalo interaction parameters, including mass, scale radius, velocity, and encounter geometry, from stellar stream kinematics. We generate particle spray simulations based on the Lagrange Cloud stripping technique, focusing on the ATLAS-Aliqa Uma stream as a test case. We train multiple NPE models across multiple observational scenarios, quantifying how kinematic completeness affects inference and forecasting constraints from upcoming surveys including LSST, 4MOST, and 10-year Gaia data. Our results demonstrate that NPE can produce accurate and well-calibrated posteriors. In the idealized case with full 6D coordinates, we achieve subhalo mass uncertainties of 15-20% for a $10^7 \, \mathrm{M_\odot}$ subhalo, with 5D coordinates (excluding radial velocities) achieving similar performance. Under realistic observational conditions, mass uncertainties range from 50% (present-day) to 20-40% (future scenarios), with comparable performance between the photometric-only LSST sample and a smaller sample that includes Gaia proper motions and 4MOST radial velocities. Most notably, we find that velocity bimodality emerges when phase space is poorly sampled, whether due to missing kinematic information or limited stellar tracers. Combining large photometric samples with targeted spectroscopic follow-up can effectively resolves this degeneracy. These results demonstrate the power of implicit likelihood inference for optimizing stellar stream observational strategies and forecasting DM subhalo constraints from upcoming surveys.

Nguyen, Tri [Northwestern U. (main); SkAI, Chicago↗

Findings from Large Bench-Scale Testing of Denitration Electrolyzers for the EDCGe Project

This report highlights the key findings and outcomes relevant to the processability of waste supernatant at Hanford using a denitration electrolyzer. The Electrosynthesis Company issued a Phase 1 report to the Savannah River National Laboratory (SRNL), summarizing the evaluation of large bench-scale denitration electrolyzers to support the electrochemical denitration and caustic generation (EDCGe) project. The Electrosynthesis Company’s report (attached as Appendix A) provides insights into the initial steps required to implement an electrolyzer system at Hanford. Phase 1 experiments focused on validating the denitration electrolyzer’s performance, operating parameters, and reaction products. The robustness of the electrochemical denitration process was demonstrated by two electrolyzer flow cell systems (a 100 cm 2 ElectroCell MP and a 150 cm 2 NESI NS01 cell), both of which achieved significant nitrate and nitrite removal (>50%) with a current efficiency of ~95% for both systems. Higher current densities (500 mA cm –2 ) improved nitrate and nitrite removal rates compared to lower current densities (333 mA cm –2 ), while maintaining a current efficiency of ~94%. The NS01 cell achieved a nitrate species removal rate of ~0.41 mol h –1 at 5 kA m –2 (equiv. to 500 mA cm –2 ). The primary reaction product was ammonia (NH 3 ), constituting 78.3–91.6% of the products (excluding OH – formation). NH 3 was predominantly retained in the catholyte liquid phase rather than being off-gassed. Additionally, the NS01 cell reported an NH 3 generation rate of ~0.36 mol h –1 at 5 kA m –2 . Other gas formation included ~7% N 2 , ~7% H 2 , and trace amounts of N 2 O. The estimated power requirement (extrapolated from the 0.015 m 2 cell data) for a full-scale denitration electrolyzer is approximated to be ~1.6 MW (DC-only) to treat 50% of nitrate and nitrite in a waste stream and generates ~2.1 kmol h –1 of NH 3 with an initial concentration of 4 M NO 3 – /NO 2 – at 300 gal h –1 . Simulated waste containing aluminate, carbonate, oxalate, and halogens exhibited no adverse effects on denitration performance. A preliminary experiment comparing alkaline anolyte (5 M NaOH) with a nickel based anode to acidic media (2 M H 2 SO 4 ) with a DSA-O 2 anode showed a lower operating voltage and generated less H 2 than the acid media. Maintaining a stable 5 M OH – concentration in the anolyte through periodic additions of caustic did not significantly impact denitration performance. This operational mode will be required for long-term experiments. All the experiments demonstrated that electrochemical denitration is a promising approach for treating nitrate and nitrite in simulated waste streams, achieving significant conversion and robustness across varying experimental conditions and electrochemical cell configurations. Lastly, the ability to generate a nearly pure NH 3 stream may prove advantageous for processing at other locations within the Hanford site.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

Model Data Archive for Manuscript Titled "Evaluation of a Coupled Surface–Subsurface Hydrologic Model Using Dense Water‑Level Sensors in a Mixed Urban–Rural Watershed"

This archive provides scripts, input files, and datasets used for the implementation and evaluation of a fully coupled surface–subsurface hydrologic model in the Neches River Basin, southeast Texas. The study uses the Advanced Terrestrial Simulator (ATS) to simulate coupled surface–subsurface hydrologic processes over a mixed urban–rural watershed and evaluates model performance using a dense network of 136 in situ water-level sensors, nine U.S. Geological Survey (USGS) stream gauges, and SSEBop-derived evapotranspiration estimates during the period October 2014–June 2024. The workflow is implemented primarily in Python 3 using the Watershed Workflow package. The Jupyter notebooks can be executed using open-source software such as Anaconda JupyterLab or Visual Studio Code. Other data files include TXT, CSV, XML, SHP, TIF, NetCDF, HDF5, and ExodusII files, which can be processed using the provided Python scripts. ATS input files are provided in XML format and can be edited using any commonly used text editor. This archive contains: *Scripts and input files used to generate the ATS model setup, including watershed discretization, mesh generation, parameter mapping, and model configuration. *Jupyter notebooks used for preprocessing observational data, evaluating streamflow, water levels, and evapotranspiration, computing performance metrics, and generating the figures presented in the manuscript. *ATS simulation outputs and processed observational datasets, including OneRain and DD6 water-level sensors, USGS streamflow observations, GIS data, and supporting spatial datasets used throughout the study.

Dense water-level sensor network↗

If We Build Them, They Will Run: Automated HPC Apps Deployment and Profiling with eBPF in Cloud

The high performance computing (HPC) community is in a period of transition. The rise of AI/ML coupled with a changing landscape of resources deems portability a new metric of performance, and methods to move between on-premises and cloud environments and assess compatibility are paramount. Here we design and test a strategy for bridging the gap between traditional HPC and Kubernetes environments – first containerizing applications, providing automated orchestration to run studies, and packaging the setup with automated means to assess performance using low overhead eXtended Berkeley Packet Filter (eBPF) programs. We first assess different designs for eBPF collection, demonstrating a tradeoff between number of programs deployed on a node and overhead added. We develop 5 low overhead eBPF programs that combine with streaming ML models to assess CPU, futex, TCP, shared memory, and file access across four different builds of an HPC application for CPU and GPU. We use eBPF data to generate insights into the possible underlying etiology of scaling issues. We then assess compatibility of a well-known benchmark, HPCG, across matrices of micro-architectures and optimization levels (217 containers across 24 instance types and over 7500 runs). We provide to the community 30 applications to deploy in our automated setup and perform a scaling study from 4 to a maximum of 256 nodes for both CPU and GPU applications. Finally, we use our gained knowledge about performance to generate compatibility artifacts that are used by a newly developed Kubernetes controller to intelligently select instance type based on optimizing a figure of merit. Along with insights to scaling in this environment with a collection of applications and templates to work from, we provide an overall strategy for approaching HPC application deployment and image selection based on compatibility in cloud.

Computer science↗

Memory-Aware External Facelist Calculation: A Data-Parallel Atomic Hash Counting Approach

Unstructured volumetric meshes serve as fundamental data representations in various scientific simulations and analyses. They play a crucial role in representing complex computational domains and are essential for important numerical techniques, such as finite element analysis. Whenever such a mesh is read from a file, streamed in-situ, or generated by algorithms, scientific visualization libraries rely on calculating the external surface of a geometry, named “external facelist”, to produce a polygonal mesh for rendering. Consequently, external facelist calculation has become one of the most widely used algorithms in the scientific visualization domain, necessitating optimal performance. In this paper, we explore relevant work on external facelist calculation algorithms in two common visualization libraries, VTK and Viskores, assess their performance and memory constraints, and introduce a novel memory-aware external facelist calculation algorithm employing an atomic hash counting approach. This algorithm fully leverages Viskores' data-parallel primitive operations, facilitating its execution across diverse many-core architectures. Our algorithm features the lowest memory footprint on the GPU and the second-lowest on the CPU among all evaluated methods, and it also delivers the fastest performance on both CPU and GPU. It has been made available under an open-source license in the VTK and Viskores visualization systems.

Tsalikis, Spiros [Kitware] (ORCID:0000000151137195↗

SoK: What does it Mean to Benchmark Database Forensics?

Relational Database Management Systems are the backbone of modern enterprises and public-sector services, and are thus frequent targets of security incidents, insider threats, and thorough regulatory audits. Consequently, databases have become key sources of digital evidence, requiring investigators to reconstruct past activity from audit logs, transaction logs, and backups. Although benchmarking frameworks such as those developed by the Transaction Processing Performance Council (TPC) are widely used to evaluate database performance, they do not capture forensic requirements such as evidentiary completeness, tamper-evidence, chain of custody, or regulatory compliance under GDPR and CCPA. This survey examines the emerging domain of forensic database benchmarking. We gathered prior research on database forensics, secure logging, and tamper-evident data structures; we analyze modern forensic-ready features in commercial and open-source systems (SQL Server Ledger, Oracle Blockchain Tables, PostgreSQL pgAudit, Db2 Audit, Aurora Database Activity Streams, Oracle Real Application Security and IBM Guardium) and assess why existing benchmarks are insufficient. We propose forensic workloads, metrics, and methodologies that incorporate adversarial stressors, deleted-record recovery, and backup analysis. We also identify open research problems and call for a community-driven forensic benchmark suite. The result is an idea for evaluating not only database performance but also forensic soundness, bridging the gap between system engineering, compliance, and digital investigations.

Lenard, Ben↗

Connecting ambient toxicity testing with community-level responses of benthic macroinvertebrates in an impacted stream in East Tennessee, USA

Single-species laboratory toxicity tests are a standard tool for evaluating potential impairment of freshwater systems; however, it remains uncertain how well they reflect community-level impacts in natural environments. This study presents a multi-decadal dataset (2005-2025) pairing ambient toxicity testing with macroinvertebrate surveys along Bear Creek on the Oak Ridge Reservation (Tennessee, USA) downstream of an industrial complex to assess the ability of laboratory tests using stream water to track community-level effects. Biannual three-brood Ceriodaphnia dubia tests from 2005 to 2025 often showed reduced reproduction at select sites. Integrating water quality data showed strong positive correlations between sublethal toxicity and specific conductance. Macroinvertebrate diversity metrics, family-level occurrence, and densities were also associated with conductance and contemporaneous C. dubia responses. Laboratory-measured sublethal toxicity was a stronger indicator of macroinvertebrate change than conductance alone, although responses varied among sites and seasons. At the site with the highest diversity, densities and richness of Ephemeroptera, Plecoptera, Trichoptera (EPT) and non-EPT taxa were significantly related to C. dubia reproduction, with greater toxicity corresponding to lower diversity. At the family level, some pollution-tolerant taxa were more prevalent and at higher densities during periods of sublethal toxicity, while some sensitive families were absent or reduced. These patterns may reflect site-specific mixtures of acute and chronic stressors, with laboratory toxicity tests more effectively capturing short-term impacts. Overall, these multi-decadal observations suggest that laboratory toxicity tests can help track water-quality changes linked to shifts in aquatic community diversity, despite variable responses reflecting the complexity of dynamic stressors in this impacted freshwater system.

Stevenson, Louise [ORNL] (ORCID:0000000349679897)↗

Nearby stellar substructures in the Galactic halo from DESI Milky Way Survey Year 1 Data Release

We report five nearby ($d_{\mathrm{helio}} < 5$ kpc) stellar substructures in the Galactic halo from a subset of 138 661 stars in the Dark Energy Spectroscopic Instrument (DESI) Milky Way Survey Year 1 Data Release. With an unsupervised clustering algorithm, HDBSCAN*, these substructures are independently identified in Integrals of Motion ($E_{\rm tot}$, $L_{\rm z}$, $\log {J_r}$, $\log {J_z}$) space and Galactocentric cylindrical velocity space ($V_{R}$, $V_{\phi }$, $V_{z}$). We associate all identified clusters with known nearby substructures (Helmi streams, M18-Cand10/MMH-1, Sequoia, Antaeus, and ED-2) previously reported in various studies. With metallicities precisely measured by DESI, we confirm that the Helmi streams, M18-Cand10, and ED-2 are chemically distinct from local halo stars. We have characterized the chemodynamic properties of each dynamic group, including their metallicity dispersions, to associate them with their progenitor types (globular cluster or dwarf galaxy). Our approach for searching substructures with HDBSCAN* reliably detects real substructures in the Galactic halo, suggesting that applying the same method can lead to the discovery of new substructures in future DESI data. With more stars from future DESI data releases and improved astrometry from the upcoming Gaia Data Release 4, we will have a more detailed blueprint of the Galactic halo, offering a significant improvement in our understanding of the formation and evolutionary history of the Milky Way Galaxy.

dynamics↗

Timeseries Photos of a Variably Inundated Stream: Umtanum Creek, Washington, United States

This dataset is associated with a broader study using game camera timeseries photos collected to evaluate stream variable inundation via changes in width (i.e. wet fraction). Four game cameras were deployed along Umtanum Creek (Washington, United States) to track changes in stream inundation over time. Drone imagery was collected at the same location on October 18, 2024 which was used to construct a digital elevation model (DEM) of the streambed topography. The associated paper and data can be found at https://doi.org/10.1016/j.envsoft.2025.106715 (Bao et al., 2025a)) and https://doi.org/10.15485/2589885 (Bao et al., 2025b), respectively. For details on how to navigate data packages generated by this project, see https://data.ess-dive.lbl.gov/portals/PNNLRiverCorridorSFA/About. In addition to this readme, this data package also includes a file-level metadata (FLMD) files that describes each file and a data dictionaries (DD) that describe all column/row headers and variable definitions. This dataset is comprised of (1) file-level metadata; (2) data dictionary; (3) readme; (4) field metadata; (5) field protocol; and (5) folders containing game camera photos. Game camera photos are organized into folders for each camera (CDL, CUL, CDR, CUR; see readme for information on camera naming) by the month photos were collected. All files are .csv, .jpg, or .pdf.

AI image segmentation↗