Search NASA⌕ Search

SEARCH · Search NASA

Results for “Data analysis methods”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Copacabana: a probabilistic membership assignment method for galaxy clusters

Cosmological analyses using galaxy clusters in optical/near-infrared photometric surveys require robust characterization of their galaxy content. Precisely determining which galaxies belong to a cluster is crucial. In this paper, we present the COlor Probabilistic Assignment of Clusters And BAyesiaN Analysis (Copacabana) algorithm. Copacabana computes membership probabilities for all galaxies within an aperture centred on the cluster using photometric redshifts, colours, and projected radial probability density functions. We use simulations to validate Copacabana and we show that it achieves up to 89 per cent membership accuracy with a mild dependence on photometric redshift uncertainties and choice of aperture size. We find that the precision of the photometric redshifts has the largest impact on the determination of the membership probabilities followed by the choice of the cluster aperture size. We also quantify how much these uncertainties in the membership probabilities affect the stellar mass–cluster mass scaling relation, a relation that directly impacts cosmology. Using the sum of the stellar masses weighted by membership probabilities (⁠μ * ⁠) as the observable, we find that Copacabana can reach an accuracy of 0.06 dex in the measurement of the scaling relation at low redshift for a Legacy Survey of Space and Time type survey. These results indicate the potential of Copacabana and μ * to be used in cosmological analyses of optically selected clusters in the future.

79 ASTRONOMY AND ASTROPHYSICS↗

On Alfvénic turbulence of solar wind streams observed by Solar Orbiter during March 2022 perihelion and their source regions

It has been recently accepted that the standard classification of the solar wind solely according to flow speed is outdated, and particular interest has been devoted to the study of the origin and evolution of so-called Alfvénic slow solar wind streams and to what extent such streams resemble or differ from fast wind. In March 2022, Solar Orbiter completed its first nominal phase perihelion passage. During this interval, it observed several Alfvénic streams, allowing for characterization of fluctuations in three slow wind intervals (AS1-AS3) and comparison with a fast wind stream (F) at almost the same heliocentric distance. This work makes use of Solar Orbiter plasma parameters from the Solar Wind Analyzer (SWA) and magnetic field measurements from the magnetometer (MAG). The magnetic connectivity to the solar sources of selected solar wind intervals was reconstructed using a ballistic extrapolation based on measured solar wind speed down to the (spherical) source surface at 2.5 R s below which a potential field extrapolation was used to map back to the Sun. The source regions were identified using SDO/AIA observations. A spectral analysis of in situ measured magnetic field and velocity fluctuations was performed to characterize correlations, Alfvénicity, normalized cross-helicity, and residual energy in the frequency domain as well as intermittency of the fluctuations and spectral energy transfer rate estimated via mixed third-order moments. A machine learning technique was used to separate proton core, proton beam, and alpha particles and to study v − b correlations for the different ion populations in order to evaluate the role played by each population in determining the Alfvénic content of solar wind fluctuations. The comparison between fast wind and Alfvénic slow wind intervals highlights the differences between the two solar wind regimes: The fast wind is characterized by larger amplitude fluctuations, and magnetic and velocity fluctuations are closer to equipartition of energy. In fact the Alfvénic slow wind streams appear to be on a spectrum of wind types, with AS1, originating from open field lines neighboring active regions and displaying similarities with the fast wind in terms of fluctuation amplitude and turbulence characteristics, but not with respect to the alpha particles and proton beams. The other two slow streams differed both in their sources as well as plasma characteristics, with AS2 coming from the expansion of a narrow coronal hole corridor and AS3 from a region straddling a pseudostreamer. The latter displayed the coldest and highest density but the slowest stream with the smallest fluctuation amplitude and greatest magnetic energy excess. It also showed the largest scatter in proton beam speeds and the greatest difference in speed between proton beam and alpha particles. This study shows how the old fast–slow solar wind dichotomy, already called into question by the observations of slower Alfvénic solar wind streams, should further be refined, as the Alfvénic slow wind, originating in different solar wind regions, show significant differences in density, temperature, and proton and alpha-particle properties in the inner heliosphere. The observations presented here provide the starting point for a better understanding of the origin and evolution of different solar wind streams as well as the evolving turbulence contained within.

magnetohydrodynamics (MHD)↗

Informed unsupervised machine learning analysis of dislocation microstructure from high-resolution differential aperture X-ray structural microscopy data

This study leverages high-resolution differential-aperture X-ray structural microscopy (DAXM) to probe the local dislocation structure in deformed 304L-stainless steel at small strain, by measuring the lattice rotation and deviatoric elastic strain with a sub-micron resolution. For a single grain in a polycrystalline specimen, the measured lattice rotation field over the measured volume exhibited a multimodal distribution while the deviatoric elastic strain showed a single-mode distribution. An unsupervised Cauchy mixture machine learning model was developed to resolve the multimodal distribution of the lattice rotation. By mapping the lattice rotation data associated with each Cauchy peak in the model back onto the measured volume, we identify contiguous regions of the crystal rotated near the average values corresponding to the peaks of the overall rotation distribution. These regions represent the grain subdivision in the microstructure. Finally, the dislocation density tensor was also computed and its norm was laid over the rotation field to detect the subgrain boundaries. This step provided a validation of the Cauchy mixture model for the analysis of the lattice rotation distribution. The current study highlights the integration of advanced X-ray microscopy techniques with data-driven analysis methods to uncover detailed microstructure scales in deformed crystals.

Machine learning; Lattice rotation; High-energy X-↗

Optimising the processing and storage of visibilities using lossy compression

The next-generation radio astronomy instruments are providing a massive increase in sensitivity and coverage, largely through increasing the number of stations in the array and the frequency span sampled. The two primary problems encountered when processing the resultant avalanche of data are the need for abundant storage and the constraints imposed by I/O, as I/O bandwidths drop significantly on cold storage. An example of this is the data deluge expected from the SKA Telescopes of more than 60 PB per day, all to be stored on the buffer filesystem. While compressing the data is an obvious solution, the impacts on the final data products are hard to predict. In this paper, we chose an error-controlled compressor – MGARD – and applied it to simulated SKA-Mid and real pathfinder visibility data, in noise-free and noise-dominated regimes. As the data have an implicit error level in the system temperature, using an error bound in compression provides a natural metric for compression. MGARD ensures the compression incurred errors adhere to the user-prescribed tolerance. To measure the degradation of images reconstructed using the lossy compressed data, we proposed a list of diagnostic measures, exploring the trade-off between these error bounds and the corresponding compression ratios, as well as the impact on science quality derived from the lossy compressed data products through a series of experiments. We studied the global and local impacts on the output images for continuum and spectral line examples. We found relative error bounds of as much as 10%, which provide compression ratios of about 20, have a limited impact on the continuum imaging as the increased noise is less than the image RMS, whereas a 1% error bound (compression ratio of 8) introduces an increase in noise of about an order of magnitude less than the image RMS. For extremely sensitive observations and for very precious data, we would recommend a 0.1% error bound with compression ratios of about 4. These have noise impacts two orders of magnitude less than the image RMS levels. At these levels, the limits are due to instabilities in the deconvolution methods. We compared the results to the alternative compression tool DYSCO, in both the impacts on the images and in the relative flexibility. MGARD provides better compression for similar error bounds and has a host of potentially powerful additional features.

Techniques: interferometric↗

Elliptic multipoles and the modeling of narrow-gap bend magnets in accelerators

We highlight the virtues of 2D elliptic-multipole field expansions in modeling the magnetic fields of narrow-aperture, straight-axis bending magnets with parallel faces, addressing the limitations of the conventional circular multipole series when the beam-orbit sagitta exceeds the magnet's vertical half-gap. The elliptic multipoles provide a convenient way to represent the field in all aspects of the magnet development (design, particle-tracking simulations, measurements). We propose a numerically robust method of data analysis to determine the elliptic (or circular) multipoles from stretched-wire measurements with the wire moving on an arbitrary path.

Venturini, Marco↗

Measurements of the z > 5 Lyman-α forest flux autocorrelation functions from the extended XQR-30 data set

We present the first observational measurements of the Lyman-α (Ly α) forest flux autocorrelation functions in ten redshift bins from 5.1 ≤ z ≤ 6.0. We use a sample of 35 quasar sightlines at z > 5.7 from the extended XQR-30 data set; these data have signal-to-noise ratios of >20 per spectral pixel. We carefully account for systematic errors in continuum reconstruction, instrumentation, and contamination by damped Ly α systems. With these measurements, we introduce software tools to generate autocorrelation function measurements from any simulation. Our measurements of the smallest bin of the autocorrelation function increase with redshift when normalizing by the mean flux, $\langle{F}\rangle$. This increase may come from decreasing $\langle{F}\rangle$ or increasing mean free path of hydrogen-ionizing photons, λmfp. Recent work has shown that the autocorrelation function from simulations at z > 5 is sensitive to λmfp, a quantity that contains vital information on the ending of reionization. For an initial comparison, we show our autocorrelation measurements with simulation models for recently measured λmfp values and find good agreements. Further work in modelling and understanding the covariance matrices of the data is necessary to get robust measurements of λmfp from this data.

79 ASTRONOMY AND ASTROPHYSICS↗

Detection of the large-scale tidal field with galaxy multiplet alignment in the DESI Y1 spectroscopic survey

We explore correlations between the orientations of small galaxy groups, or ‘multiplets’, and the large-scale gravitational tidal field. Using data from the Dark Energy Spectroscopic Instrument (DESI) Y1 survey, we detect the intrinsic alignment (IA) of multiplets to the galaxy-traced matter field out to separations of $100\,h^{-1}$ Mpc. Unlike traditional IA measurements of individual galaxies, this estimator is not limited by imaging of galaxy shapes and allows for direct IA detection beyond redshift $z=1$. Multiplet alignment is a form of higher order clustering, for which the scale-dependence traces the underlying tidal field and amplitude is a result of small-scale ($\lt 1h^{-1}$ Mpc) dynamics. Within samples of bright galaxies, luminous red galaxies (LRG) and emission-line galaxies, we find similar scale-dependence regardless of intrinsic luminosity or colour. This is promising for measuring tidal alignment in galaxy samples that typically display no IA. DESI’s LRG mock galaxy catalogues created from the A BACUS S UMMIT N -body simulations produce a similar alignment signal, though with a 33 per cent lower amplitude at all scales. An analytic model using a non-linear power spectrum (NLA) only matches the signal down to 20 $h^{-1}$ Mpc. Our detection demonstrates that galaxy clustering in the non-linear regime of structure formation preserves an interpretable memory of the large-scale tidal field. Multiplet alignment complements traditional two-point measurements by retaining directional information imprinted by tidal forces, and contains additional line-of-sight information compared to weak lensing. This is a more effective estimator than the alignment of individual galaxies in dense, blue, or faint galaxy samples.

79 ASTRONOMY AND ASTROPHYSICS↗

Galaxy-multiplet clustering from DESI DR2

We present an efficient estimator for higher-order galaxy clustering using small groups of nearby galaxies, or multiplets. Using the Luminous Red Galaxy (LRG) sample from the Dark Energy Spectroscopic Instrument (DESI) Data Release 2, we identify galaxy multiplets as discrete objects and measure their cross-correlations with the general galaxy field. Our results show that the multiplets exhibit stronger clustering bias as they trace more massive dark matter halos than individual galaxies. When comparing the observed clustering statistics with the mock catalogs generated from the N-body simulation AbacusSummit, we find that the mocks underpredict multiplet clustering despite reproducing the galaxy two-point auto-correlation reasonably well. This discrepancy indicates that the standard Halo Occupation Distribution (HOD) model is insufficient to describe the properties of galaxy multiplets, revealing the greater constraining power of this higher-order statistic on galaxy-halo connection and the possibility that multiplets are specific to additional assembly bias. We demonstrate that incorporating secondary biases into the HOD model improves agreement with the observed multiplet statistics, specifically by allowing galaxies to preferentially occupy halos in denser environments. Our results highlight the potential of utilizing multiplet clustering, beyond traditional two-point correlation measurements, to break degeneracies in models describing the galaxy-dark matter connection.

cosmology↗

Field Validation of MVA Technology for Offshore CCS: Novel Ultra-High-Resolution 3D Marine Seismic Technology (P-Cable) (Final Report)

The objectives of the proposed study were to deploy and validate a specific monitoring technology, high-resolution 3D marine seismic (HR3D), appropriate for large-demonstration and commercial-scale offshore CCS sites. The project accomplished successful acquisition two HR3D seismic surveys. The first HR3D dataset was over the offshore injection site of the Tomakomai, Japan integrated pilot CCS project, which at the time of survey acquisition was actively injecting CO 2 . The first survey also represented a successful international collaboration between the DOE NETL program and Japan’s national CCS program and was the first successful acquisition and use of HR3D over an active CO 2 injection site (Meckel, Feng et al. 2019). The Tomakomai HR3D survey successfully tested a novel 4-streamer HR3D system array in which, for the first time, no cross-cable (aka “P-Cable”) was utilized and only four GeoEel streamers were used instead of the standard 12-streamer configuration. Consequently, this was not, strictly speaking, a deployment of the “P-Cable” system of (Planke and Berndt 2004) but rather a modified version, thereof, and it is the first known demonstration of the modified system configuration. One very positive outcome from the Japanese collaboration earlier in the project was the ability to learn from the Japanese how they used tail buoys with GPS to determine the position of the seismic source and receivers in time and space. Based on that experience, GCCC designed and built six GPS receivers that could be used to position the streamer receivers and the seismic source via tail buoys. A fundamental advance that was made on the original design, was the ability to directly power the tail buoy GPS units and transfer data through the streamers (i.e., vs. the batteries used at Tomakomai). The bulkiness of the GPS batteries caused drag and episodic surging of the buoys, which affected data quality by lifting up the tail end of the streamers so the receivers were not at the same depth. The units were tested onshore for accuracy and functionality, and the design was subsequently and successfully tested in marine acquisition mode during the SLP survey acquisition. The marine acquisition test and survey satisfied Subtasks 2.2.2, Novel Positioning Technology Selection and Subtask 2.2.3, Novel Positioning Technology Deployment. Results of the novel positioning technology selection (Subtask 2.2.2) were considered successful and will be incorporated in future HR3D seismic acquisition projects to reduce costs, improve deployment safety at sea, and integrate both seismic and data recording via a single data transfer through the streamers to the recording system. The project also established a permitting process through NETL NEPA compliance, which included an Environmental Assessment in a marine setting and is required for conducting these types of surveys using Federal funding. The permitting process charted a “boilerplate,” which can allow future surveys related to other funded projects to move forward more expeditiously. Future improvements that could be considered are more robust seals on the GPS module and stronger materials (especially joints) on tail buoy fabrication. These would increase fixed costs, but would be advisable and probably more economic long-term if multiple HR3D surveys are planned. Project Accomplishments include: • Pre-survey Sensitivity Study • Marine geochemistry methods and data analysis • Successful HR3D seismic dataset acquired @ Tomakomai active CO 2 injection marine site • Developed advanced seismic processing techniques • No NRMS anomalies detected in overburden; Demonstration of containment • Repeatability study • Second survey collected @ San Luis Pass, TX • 4D application using positioning techniques developed in the project for monitoring were successful

3D seismic GPS positioning↗

Developing ML/AI Methods for High-Throughput Characterization of Multiple-Sensor Streams of Tokamak Dynamics for High-Speed Control (Final Report)

This project evaluated and developed new mathematical and algorithmic techniques capable of handling (in real-time) the growing amounts of data generated by modern fusion research. While existing numerical linear algebra (NLA) methods provide the backbone to classical data analysis and algorithms, these methods fundamentally do not port to distributed architectures nor do they allow low-latency data reduction for control. Motivated by the needs for modern fusion reactors, this project explored and implemented new numerical methods to characterize plasma dynamics, respond in real-time to discharge evolution, and to process massive-scale data accurately and rapidly more fully. This project links expertise in multiple-sensor diagnostics of tokamak plasma dynamics from Columbia University’s Plasma Physics Laboratory with expertise in massive-scale data reduction and extreme data control algorithms at Columbia University’s Data Science Institute. This interdisciplinary project (i) applied machine learning methods, (ii) implemented a properly-trained neural-network for very fast processing of high-speed plasma videography, and (ii) developed the applied mathematical methods, based on randomized-NLA (rNLA) routines, for data analysis, reduction, and real-time control. The Columbia University High Beta Tokamak-Extended Pulse (HBT-EP) facility provided data to test new algorithms and partnership with Columbia University's Data Sciences Institute evaluated the broader use of new algorithms for many challenging control applications.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Holistic energy analysis method for thermal management architectures of data centers

Modern high-performance computing (HPC) data centers (DCs), particularly those supporting energy-intensive artificial intelligence (AI) workloads, face escalating thermal management challenges that degrade performance through thermal throttling and drive up cooling power consumption and operational costs. To address this challenge, many have developed a wide variety of thermal management solutions (single-phase, two-phase, direct, indirect, hybrid, and more) which attempt to cool HPC DCs effectively while attempting to minimize overall system power consumption. However, the analysis of these solutions and methods to effectively compare one with another is lacking. Overall power usage effectiveness (PUE) and total-power usage effectiveness (TUE) provide a metric to quantify power consumption but fail to identify components in the system which require further optimization. To address this, we propose a holistic analytical framework – the waterfall diagram (WFD) – which leverages a waterfall chart methodology, offering a comprehensive visualization of both the thermal management system loop and heat flow pathways from individual server components to the outdoor ambient. Use of the WFD enables graphical estimations of power efficiency and cooling performance across each component of a DC cooling system and complements Sankey-style energy flow visualizations by additionally resolving stage-wise temperature changes and incremental TUE contributions. The framework is used in conjunction with simulation-based approaches, to conduct a detailed pressure drop and flow distribution analysis aimed at identifying the optimal coolant distribution architecture for a single-phase direct-to-chip water-cooled DC, which serves as the baseline for subsequent WFD analysis. Among the evaluated architectures, the 3 U modular coolant distribution architecture is found to demonstrate the best performance, considering minimal pressure drop and uniform flow distribution. In addition, TUE is calculated for each cooling loop component based on its associated pressure drop and corresponding pumping power, which are integrated into the WFD. This correlation between TUE and local temperature offers immediate insight into the power efficiency and thermal performance contributions of individual components, facilitating further development and optimization. Examples of WFD applications are presented under varying thermal loads and ambient conditions, demonstrating reasonable cooling strategies. Notably, the 3 U modular architecture maintains a consistent chip case temperature of 85°C, achieving a TUE of 1.016 at ambient temperature of 47°C, and a TUE of 1.026 at ambient temperature of 52°C. The WFD methodology provides an efficient, holistic, and streamlined framework for DC thermal management architecture assessment and enables design optimization which is important for addressing the thermal-fluidic energy challenges of current and next-generation DCs.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

From Machine Learning to Machine Reasoning: A Model-based Approach to Analyze Equipment Reliability Data

In current nuclear power plants (NPPs) a large amount of condition-based data which can be used to assess and monitor component health and performance. Assessing component health from such data can be performed with a large variety of methods. While the analysis of numeric data can be performed with several methods, the extraction of information from textual data remains a challenge. Currently employed natural language processing (NLP) methods do not really provide quantitative information that might be contained in IRs. In addition, the integration of numeric and textual data to identify possible causal relationships between data elements is still an unresolved challenge. This paper presents an approach to extract information from textual (e.g., incident or maintenance reports) and numeric data that relies on model based system engineer (MBSE) models. MBSE are diagrams designed to represent system and component dependencies (from both a form and functional point of view). In our approach, MBSE models emulate system engineer knowledge about component/system architecture. NLP methods are employed to perform syntactic and semantic analyses. Syntactic analysis analyzes the grammatical structure of a sentence while semantic analysis is designed to analyze the logic structure of a sentence. An innovative element of our approach is that semantic analysis uses MBSE models to identify links between textual elements. Similarly, numeric data is directly linked to elements of the MBSE models in order to map which functions are being monitored.

97 - MATHEMATICS AND COMPUTING↗

Preliminary Study on Fine-Grained Power and Energy Measurements on Grace Hopper GH200 with Open-Source Performance Tools

The increasing adoption of tightly integrated, heterogeneous architectures, combined with the slowdown of Moore’s law, has made application power and energy-driven optimizations critical to efficiently use high-performance computing systems. This paper introduces a newly developed open-source toolkit that seamlessly integrates the Linux real-time hardware monitoring program hwmon with the Performance Application Programming Interface and the Score-P performance measurement system, thereby enabling fine-grained power and energy measurements for high-performance computing applications. Our primary target platform is the Wombat test bed, which is a system based on the NVIDIA GH200 superchip. The toolkit can capture transient power peaks with high temporal resolution (50 ms) and, thanks to Score-P integration, can map power metrics to specific code regions, thereby providing actionable information on power-intensive operations and inefficiencies. The toolkit also provides a holistic view of both the power and the energy consumption of the entire GH200 superchip by covering all major components: the Grace CPU, the Hopper GPU, and the I/O subsystem. Experiments that use Locally Self-consistent Multiple Scattering, which is an application for first-principles calculations of materials developed at Oak Ridge National Laboratory, have demonstrated the tool’s ability to identify transient power spikes and uncover opportunities for energy-aware optimizations. Additionally, we introduce a Python-based utility for converting Open Trace Format 2 traces to Parquet format, thus enabling advanced data analysis for numerical integration methods applied to power data for accurate energy profiling.

Hernandez Mendoza, Oscar [ORNL] (ORCID:00000002538↗

Author Correction: US oil and gas system emissions from nearly one million aerial site measurements

Correction to: Naturehttps://doi.org/10.1038/s41586-024-07117-5 Published online 13 March 2024 In the version of the article initially published, several errors were present and have been corrected in the HTML and PDF versions of the article and Supplementary Information. The main results, conclusions, and our interpretations of the data remain unchanged. See the new Supplementary Information Section S15 for a more detailed description of the errors corrected and the resulting effects on the analysis. Data processing and methods corrections Overflight count correction: We previously used pre-computed source coverage data for some Carbon Mapper campaigns that was computed differently than was required for our analysis. We have re-computed Carbon Mapper source coverage based on flightline polygons and source coordinates. Transition point computation, well sites: The updated version now correctly compares the cumulative emissions distribution of simulated well site emissions with that of aerially detected sources (rather than plumes) when computing the transition point. Transition point computation, midstream: Additionally, the transition point calculation has been corrected to exclude aerially detected midstream emissions below the transition point, which was previously leading to double counting of these emissions. This error was not present for upstream (well site) emissions. Calculation errors Unit error: We corrected a specific unit conversion error affecting well site emissions in the Kairos Fort Worth dataset. Across all datasets, we also correct the conversion factor for converting from standard volume to mass for midstream emissions. Sorting error: We correct code that was applying incorrect sorting when computing correction factors to account for partial detection at well sites. Small typographical corrections were made in Fig. 1b and SI Section S4.1. Data processing and methods corrections Overflight count correction: We previously used pre-computed source coverage data for some Carbon Mapper campaigns that was computed differently than was required for our analysis. We have re-computed Carbon Mapper source coverage based on flightline polygons and source coordinates. Transition point computation, well sites: The updated version now correctly compares the cumulative emissions distribution of simulated well site emissions with that of aerially detected sources (rather than plumes) when computing the transition point. Transition point computation, midstream: Additionally, the transition point calculation has been corrected to exclude aerially detected midstream emissions below the transition point, which was previously leading to double counting of these emissions. This error was not present for upstream (well site) emissions. Calculation errors Unit error: We corrected a specific unit conversion error affecting well site emissions in the Kairos Fort Worth dataset. Across all datasets, we also correct the conversion factor for converting from standard volume to mass for midstream emissions. Sorting error: We correct code that was applying incorrect sorting when computing correction factors to account for partial detection at well sites. Small typographical corrections were made in Fig. 1b and SI Section S4.1. The following practices may help researchers conducting similar analyses avoid making similar errors: 1, Clear, accessible documentation explaining the interpretation of all columns in data input tables and all internal variables within the model, 2, Simple cross-check calculations computed before and after unit conversions.

Sherwin, Evan D↗

Understanding the structural and morphological effects of synthesis route on NpO 2

The availability of actinide standard materials for use in nuclear safeguard applications is critical, as is thorough characterization thereof. Although accurate trace element compositions and isotopic considerations are paramount for deployment of reference standards, structural characterization is also essential towards accurately describing the chemical form and potential matrix effects in candidate materials. Here, to this end, samples of NpO 2 were synthesized via a direct denitration (DD) method and probed with powder X-ray diffraction (PXRD), Raman spectroscopy, and scanning electron microscopy (SEM) for structural and morphological characterization and comparison with NpO 2 materials produced via modified direct denitration (MDD). PXRD confirmed the bulk identity of NpO 2 , and no additional phases were identified using this method. Analysis of Raman data collected using a 532 nm excitation wavelength indicates that samples are mostly phase pure; however, some variability in spectral features is observed. Analysis of additional spectroscopic data collected with a 785 nm excitation wavelength revealed variability in the relative intensity of spectral features. Raman spectroscopy indicates that the sample is primarily NpO 2 ; however, additional signals indicate possible structural disorder, oxidized species, or potential contributions from other Np phases. To further investigate the possibility of additional phase contributions within the sample of NpO 2 , Raman spectroscopic mapping was employed to examine the homogeneity of the sample produced via DD. From this analysis, we determined that despite variability in the intensity of Raman-active vibrational modes, consistent spectra are obtained throughout the area of the sample investigated. SEM images show aggregates with variable sizes and shapes, with rounded, primary particles possessing an average diameter of approximately 100 nm. Comparison of the results of these multimodal analyses to the literature indicates that the crystal chemical, spectroscopic, and microstructural properties of NpO 2 vary based on synthesis method, even if X-ray diffraction data indicate that the bulk phase is NpO 2 .

Direct denitration↗

Juvenile Salmon and Their Habitats in the Columbia River Estuary: A Review and Synthesis of Knowledge Development 2000–2025

[This is a 90% discussion draft.] This is the third Synthesis Memorandum funded by the U.S. Army Corps of Engineers and developed for the Columbia Estuary Ecosystem Restoration Program (CEERP) on the topic of habitat restoration in the Columbia River Estuary (CRE) from Bonneville Dam to the river mouth. While the first two were developed by PNNL and NOAA without the benefit of stakeholder participation, for the current memo, two key activities were initiated: (1) review, by the Expert Regional Technical Group (ERTG), of status and trends monitoring and action effectiveness monitoring funded by CEERP, and (2) a workshop including representatives of the Bonneville Power Administration and the U.S. Army Corps of Engineers (the action agencies [AAs]), the National Oceanic and Atmospheric Administration (NOAA), major research agencies contributing to CEERP, and sponsors who implement CEERP restoration actions. A systematic literature review was conducted using ClarivateTM Web of ScienceTM database. The topics of interest for CRE relevant research included salmon ecology, physical processes, and wetland habitats, and therefore required the use of broad search terms. Our final search criteria included a combination of Boolean operators and an approach to combine different sets of search terms. The final search result yielded 669 records. The records were classified by groups and assigned to the relevant disciplinary expert for review. The review identified substantive advances in understanding the provision of salmon habitat functions through spatiotemporally dynamic physical and ecological processes, and the use of CRE habitats by numerous stocks of juvenile salmon. It also uncovered heretofore unincorporated historical documentation of riparian habitats across the CRE. The characterization of the structural components of floodplain habitat including plant associations and channel networks has advanced considerably, together with the understanding of seasonal changes and long-term trends. The relative influence of salmon-habitat location in the CRE as compared with temporal factors, mainly season, has been well described, which affects the prioritization of restoration. Stressors on the ecosystem and fish, and the drivers of these stressors, have been more carefully elucidated and predictive models are in various stages of development. The vision, aims, and design of restoration projects have advanced together with methods of data collection, analysis, and modeling that have seen substantial improvements. Experiments intended to inform the design of restoration projects are underway or have been completed. An important outstanding area of research that has lagged behind the advances in fundamental understanding of the ecosystem and salmon habitat functions remains the peer-reviewed documentation of the outcomes of restoration for both habitats and fish functions.

estuary↗

Time series methods for the analysis of soundscapes and other cyclical ecological data

Biodiversity monitoring has entered an era of ‘big data’, exemplified by a near-continuous collection of sounds, images, chemical and other signals from organisms in diverse ecosystems. Such data streams have the potential to help identify new threats, assess the effectiveness of conservation interventions, as well as generate new ecological insights. However, appropriate analytical methods are often still missing, particularly with respect to characterizing cyclical temporal patterns. Here, we present a framework for characterizing and analysing ecological responses that represent nonstationary, complex temporal patterns and demonstrate the value of using Fourier transforms to decorrelate continuous data points. In our example, we use a framework based on three approaches (spectral analysis, magnitude squared coherence, and principal component analysis) to characterize differences in tropical forest soundscapes within and across sites and seasons in Gabon. By reconstructing the underlying, cyclic behaviour of the soundscape for each site, we show how one can identify circadian patterns in acoustic activity. Soundscapes in the dry season had a complex diel cycle, requiring multiple harmonics to represent daily variation, while in the wet season there was less variance attributable to the daily cyclic patterns. Our framework can be applied to most continuous, or near-continuous ecological data collected at a fine temporal resolution, allowing ecologists to explore patterns of temporal autocorrelation at multiple levels for biologically meaningful trends. Such methods will become indispensable as biological big data are used to understand the impact of anthropogenic pressures on biodiversity and to inform efforts to mitigate them.

54 ENVIRONMENTAL SCIENCES↗

Common Column Identification for Table Similarity Detection in Electrified Transportation Data Lakes

Electrified transportation often requires researchers and operators to interact with datasets from a wide range of sources and disciplines, such as transportation, power systems, public health, policies, and regulations. These datasets vary in quality and format, making it difficult to understand, preprocess, and identify key columns representing real-world entities or values for indexing and joining, which can negatively impact downstream analysis and operation. Existing solutions are limited, requiring extensive manual customization or data expertise to utilize. In this article, we propose a multi-layered approach to automatically identify key columns to expedite preprocessing and aid in analysis of electrified transportation data. Our method leverages a dynamic ontology to identify common fields and an information theory-based strategy for edge cases that are difficult to generalize. Evaluations on a number of datasets from data.gov and kaggle.com show improved performance of our methods over several baseline techniques, and our ablation analyses illustrate the efficacy of individual components of our method. Our case studies also demonstrate that our methods have the potential to improve analysis of electrified transportation data and aid in automatic integration of such datasets.

33 ADVANCED PROPULSION SYSTEMS↗