Search NASA⌕ Search

SEARCH · Search NASA

Results for “sample processing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 199 records · Page 11

WHONDRS 2016 Sediment Organic Matter Characterization Data from Streams across HJ Andrews Experimental Forest, Oregon

This dataset supports a broader synoptic effort to map morphological, hydrological, chemical, and biological conditions across a fifth-order mountain stream network. Samples were generated through a collaborative synoptic sampling effort in 2016. The dataset provides sediment Fourier Transform Ion Cyclotron Resonance Mass Spectrometry (FTICR-MS) from 60 sites across the HJ Andrews Experimental Forest, Oregon (https://andrewsforest.oregonstate.edu). Related data were collected as part of the event and were published separately in collaboration with other team members. The data are available at http://www.hydroshare.org/resource/ea6c0832885a46c3939e7bb22e48e754 and are described within https://doi.org/10.5194/essd-11-1567-2019 (Ward et al., 2019). The hydroshare data package contains processed FTICR-MS data from the samples included in this data package. The data were processed via Formultitude (previously called Formularity; https://github.com/PNNL-Comp-Mass-Spec/Formultitude). However, we have re-processed the data using Core-MS and included it in this data package. Additional related data collected in 2025 from a similar effort can be found at https://data.ess-dive.lbl.gov/datasets/doi:10.15485/3023310 and http://www.hydroshare.org/resource/b274c4a234bf4b12b7cb8a54a696c629. For details on how to navigate data packages generated by this project, see https://data.ess-dive.lbl.gov/portals/PNNLRiverCorridorSFA/About. In addition to a readme, this data package also includes a file-level metadata (FLMD) file that describes each file and a data dictionary (DD) that describes all column/row headers and variable definitions. This dataset is comprised of (1) a folder of sample data; (2) data dictionary; (3) file-level metadata; (4); (5) coordinates; and (6) readme. The sample data subfolder contains 12 Tesla (12T) FTICR-MS data. This folder contains the processed data and three subfolders, one containing the .xml files, one containing the CoreMS output files, and the other containing instructions and scripts for processing the files in CoreMS (https://github.com/EMSL-Computing/CoreMS). All files are .csv, .pdf, .R, .xml, .Rmd, .py, .cal, or .json.

Biogeochemistry↗

Multi-physics melt pool modeling and process optimization for laser direct energy deposition of Nb-based refractory C103: Defect formation, geometric precision, and process mapping

Recent developments in additive manufacturing (AM) technology have reignited interest in the fabrication of the Nb-based refractory C103 alloy offering solutions to the challenges posed by traditional manufacturing methods. However, the limited numerical and experimental studies on laser direct energy deposition (DED) of C103 have hindered the understanding of the relationships between process parameters and build quality. This has made it challenging to consistently produce parts with the desired quality and microstructure suitable for critical applications. In this study, we focus on optimizing the laser DED process for C103 by employing a hybrid approach that combines experimental techniques and computational fluid dynamics (CFD). This approach facilitates the development of process maps for defect detection and geometric precision. To achieve this, multi-layer C103 samples were fabricated using laser DED under various process parameters, enabling the creation of a process map for defect detection. Additionally, a multi-physics, multiphase simulation framework was developed within a high-performance computing (HPC) environment to establish process maps for geometric precision. Using these process maps, printability windows were identified for achieving both the desired geometric accuracy and defect-free prints. It was observed that prints with a power-to-velocity (P/V) ratio close to unity resulted in defect-free outcomes. This study provides a foundation for reducing design lead time and rejected parts, ultimately optimizing the laser DED process for C103.

Defect formation and geometric precision↗

Out-of-distribution detection with non-parametric density estimation for models predicting processing history of uranium ore concentrates

The rapid advancement in machine learning (ML) and computer vision (CV) coincides with the growth of interest in deploying these ML/CV models in numerous fields from medicine to social science. Similar to those areas, we have witnessed a great number of works in materials science employing ML/CV models – neural networks in particular – in their studies in recent years. These models have proven to obtain accurate performance in various tasks. However, these models struggle to attain a similar performance when encountering test samples coming from a distribution that is different from the training set. More importantly, they fail without providing any warning to the users. Therefore, we propose a framework for detecting out-of-distribution (OOD) samples to alert users when a human intervention might be necessary in this work. Specifically, we explore the use of a non-parametric density estimation method to detect OOD samples. Here, we assess OOD detection capability of the proposed framework on ML models developed for categorizing precipitation routes of U 3 O 8 when encountering OOD datasets that contain samples (1) undergone different imaging acquisition process, (2) undergone different material synthesis process, and (3) different materials than ID set. Through those experiments, we achieve an average area under the receiver operating characteristic (AUROC) of at least 91% on average in detecting OOD samples. With minimal overhead cost and superior performance, the proposed framework enables a reliable and safe system when deploying in real-world scenarios.

Convolutional neural networks↗

Ubiquitous short-range order in multi-principal element alloys

Recent research in multi-principal element alloys (MPEAs) has increasingly focused on the role of short-range order (SRO) on material performance. However, the mechanisms of SRO formation and its precise control remain elusive, limiting the progress of SRO engineering. Here, leveraging advanced additive manufacturing techniques that produce samples with a wide range of cooling rates (up to 10 7 K s –1 ) and an enhanced semi-quantitative electron microscopy method, we characterize SRO in three CoCrNi-based face-centered-cubic (FCC) MPEAs. Surprisingly, irrespective of the processing and thermal treatment history, all samples exhibit similar levels of SRO. Atomistic simulations reveal that during solidification, prevalent local chemical order arises in the liquid-solid interface (solidification front) even under the extreme cooling rate of 10 11 K s –1 . This phenomenon stems from the swift atomic diffusion in the supercooled liquid, which matches or even surpasses the rate of solidification. Therefore, SRO is an inherent characteristic of most FCC MPEAs, insensitive to variations in cooling rates and even annealing treatments typically available in experiments.

36 MATERIALS SCIENCE↗

Cation Data for the East River Watershed, Colorado (2014-2025)

This data package contains mean values for cation concentration for water samples taken from the East River Watershed in Colorado. Inductively coupled plasma mass spectrometry (ICP-MS) has been used to measure the concentrations of elements of interest simultaneously for the East River Watershed, Colorado groundwater and surface water samples to inform insights on the biogeochemistry processes within the watershed. The East River is part of the Watershed Function Scientific Focus Area (WFSFA) located in the Upper Colorado River Basin, United States. For samples collected prior to 06-16-2021, the instrumentation, Elan DRC II, PerkinElmer SCIEX, automatically switches among the three models necessary to analyze all 37 elements. These 37 elements include: (1) Lithium (Li), Beryllium (Be), Boron (B), Sodium (Na), Magnesium (Mg), Aluminium (Al), Silicon (Si), Phosphorus (P), Titanium (Ti), Cobalt (Co), Nickel (Ni), Copper (Cu), Zinc (Zn), Germanium (Ge), Arsenic (As), Rubidium (Rb), Strontium (Sr), Zirconium (Zr), Molybdenum (Mo), Silver (Ag), Cadmium (Cd), Tin (Sn), Antimony (Sb), Caesium (Cs), Barium (Ba), Europium (Eu), Lead (Pb), Thorium (Th), Uranium (U) using standard model, argon Ar as reaction gas, (2) Potassium (K), Calcium (Ca), Vanadium (V), Chromium (Cr), Manganese (Mn), Iron (Fe) using dynamic reaction cell (DRC) model, ammonia NH3 as reaction gas, and (3) Phosphorus (P) and Selenium (Se) using DRC model, oxygen O2 as reaction gas. Note for the samples with higher concentrations of chloride (Cl-), asenic (As) concentrations were analysed with DRC model (oxygen O2 as reaction gas) to avoid the interference of chloride. For samples collected on and after 06-16-2021, an advanced Agilent 8900 triple quadrupole inductively coupled plasma mass spectrometry system (Agilent 8900 QQQ ICP-MS, Agilent Technologies) has been used to measure the concentrations of interested 36 elements simultaneously for environmental samples, including (1) Lithium (Li), Beryllium (Be) and Boron (B) using standard no gas mode, (2) Sodium (Na), Magnesium (Mg), Aluminium (Al) Phosphorus (P), Potassium (K), Chromium (Cr), Manganese (Mn), Iron (Fe), Cobalt (Co), Nickel (Ni), Copper (Cu), Zinc (Zn), Germanium (Ge), Arsenic (As), Rubidium (Rb), Strontium (Sr), Zirconium (Zr), Molybdenum (Mo), Silver (Ag), Cadmium (Cd), Tin (Sn), Antimony (Sb), Cesium (Cs), Barium (Ba), Europium (Eu), Lead (Pb), Thorium (Th) and Uranium (U) using standard helium (He) collision mode, (3) Titanium (Ti) and Vanadium (V) using high Energy (HEHe) helium (He) collision mode, and (4) Silicon (Si), Calcium (Ca) and Selenium (Se) using standard H2 reaction mode. All samples were prepared/diluted with 2% (v/v) ultrapure nitric acid in Milli-Q water (18.2 mega ohm-cm), and analyzed under a rigorous quality assurance and quality control (QA/QC) process. This data package contains (1) a zip file (cation_data_2014_2025.zip) containing a total of 5,849 files: 5.848 data files of cation data from across the Lawrence Berkeley National Laboratory (LBNL) Watershed Function Scientific Focus Area (SFA) which is reported in .csv files per location and a locations.csv (1 file) with latitude and longitude for each location; (2) a file-level metadata (v6_20260901_flmd.csv) file that lists each file contained in the dataset with associated metadata; (3) a data dictionary (v6_20260901_dd.csv) file that contains terms/column_headers used throughout the files along with a definition, units, and data type; (4) PDF and docx files for the detemination of Method Detection Limits (MDLs) for ICP-MS PerkinElmer DRC II instrumentation (Detemination_of_Method_Detection_Limits__MDLs__for_ICP_MS__PerkinElmer_Elan_DRC_II__LBL_Bldg74_Lab214D) for samples before November 2021; (5) PDF and docx files for the determination of MDLs for ICP-MS Agilent 8900 QQQ instrumentation (ICP_MS_Analysis_detection_limits_and_QA_QC_WenmingDong_updated_2026-08-06) for samples November 2021 and onward. Missing values within the anion data files are noted as either "-9999" or "0.0" for not detectable (N.D.) data. There are a total of 113 locations containing cation data. Update on 2021-04-11: Added Detemination of Method Detection Limits (MDLs) for ICP-MS document, which can be accessed as a PDF or with Microsoft Word. Update on 2022-06-10: versioned updates to this dataset was made along with these changes: (1) updated cation data for all locations up to 2021-12-31, (2) removal of units from column headers in datafiles, (3) added row underneath headers to contain units of variables, (4) removed suffix and prefix on two variables (“aqberylliumion_asberyllium” and “aqlithiumion_aslithium”), (5) added -9999 for empty numerical cells, and (6) the addition of the file-level metadata (flmd.csv) and data dictionary (dd.csv) were added to comply with the File-Level Metadata Reporting Format. Update on 2022-09-09: Updates were made to reporting format specific files (file-level metadata and data dictionary) to correct swapped file names, add additional details on metadata descriptions on both files, add a header_row column to enable parsing, and add version number and date to file names (v2_20220909_flmd.csv and v2_20220909_dd.csv). Update on 2023-08-08: Updates were made to both the data files and reporting format specific files. New available anion data was added, up until 2023-01-05. The file level metadata and data dictionary files were updated to reflect the additional data added. Update on 2024-03-11: Updates were made to both the data files and reporting format specific files. New available anion data was added, up until 2023-10-16. Further, revisions to the data files were made to remove incorrect data points (from 1970 and 2001). The reporting format specific files were updated to reflect the additional data added. Updated versions of the PDF and docx files for determination of MDLs for ICP-MS data were added to this dataset for samples starting in November 2021. Update on 2025-05-15: Updates were made to both the data files and reporting format specific files. New available cation data was added, up until the end of WY2024 (September 30, 2024). International Generic Sample Numbers (IGSNs), when registered, were added to the data files. The reporting format specific files were updated to reflect the additional data added. Update on 2026-09-01: Updates were made to both the data files and reporting format specific files. New available cation data was added, up until the end of WY2025 (September 30, 2025). Updated versions, as of 2026-08-06, of the PDF and docx files for determination of MDLs for ICP-MS data were added to this dataset for samples starting in November 2021.

54 ENVIRONMENTAL SCIENCES↗

The R -Process Alliance: Exploring the cosmic scatter among ten r -process sites with stellar abundances

Context. The astrophysical origin of the rapid neutron-capture process (r-process), responsible for producing roughly half of the elements heavier than iron, remains uncertain. Detailed chemical signatures from the oldest, most metal-poor stars, which act as fossil records of the earliest nucleosynthesis events, can be used to identify the dominant r-process sites. Aims. We present a homogeneous chemical abundance analysis of ten r-process element-enhanced stars. These old and metal-poor stars are strongly enriched in r-process elements with minimal contamination from other nucleosynthetic sources. By focusing on this chemically pure sample, we aim to investigate intrinsic variations in the r-process abundance patterns and explore their implications for the nature and potential diversity of r-process sites. Methods. We performed a detailed chemical abundance analysis of high-resolution, high-signal-to-noise spectra. For each star, we inspected over 1400 individual absorption lines using a combination of equivalent width measurements and spectral synthesis. The analysis was conducted under the assumption of 1D local thermodynamic equilibrium and employing the MOOG radiative transfer code. Results. We derived abundances for 54 chemical species, including 29 neutron-capture (n-capture) elements, covering the full mass range of the r-process abundance pattern. A kinematic analysis reveals that stars likely originated from ten kinematically distinct systems. Based on this assumption, we used the sample to probe the maximum variation expected from ten independent r-process nucleosynthesis events and computed the intrinsic dispersion of each element relative to Zr and Eu for the light and heavy r-process elements, respectively. This exercise resulted in a remarkably low cosmic scatter across the ten r-process sites enriching these stars; for the rare earth and third peak elements, for example, we find σ [La/Eu] = 0.08 and σ [Os/Eu] = 0.11 dex, while the scatter between light and heavy elements, σ [Zr/Eu] , is slightly higher at 0.18 dex. Conclusions. The elemental abundance patterns across the ten independent r-process sites show remarkably small cosmic dispersions. This minimal dispersion suggests a high degree of uniformity in r-process yields across diverse astrophysical environments.

Astronomy and AstroPhysics↗

Design of an 8-channel 40 GS/s 20 mW/Ch waveform sampling ASIC in 65 nm CMOS

One picosecond timing resolution is the entry point to signature based searches relying on secondary/tertiary vertices and particle identification. We describe PSEC5, an 8-channel 40 GS/s waveform-sampling ASIC in TSMC 65 nm process targetting one picosecond resolution at 20 mW power per channel. Each channel consists of four fast and one slow switched capacitor arrays (SCA), allowing for picosecond time resolution combined with a long effective buffer. Each fast SCA is 1.6 ns long and has a nominal sampling rate of 40 GS/s. The slow SCA is 204.8 ns long and samples at 5 GS/s. Recording of the analog data for each channel is triggered by a fast discriminator capable of multiple triggering during the window of the slow SCA. To achieve a large dynamic range, low leakage, and high bandwidth, the SCA sampling switches are implemented as 2.5 V nMOSFETs controlled by 1.2 V shift registers. Stored analog data are digitized by an external ADC at 10 bits or better. Specifications on operational parameters include a 4 GHz analog bandwidth and a dead time of 20 microseconds, corresponding to a 50 kHz readout rate, determined by the choice of the external ADC. PSEC5 has been submitted for fabrication.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Quantifying Phospholipids in Organic Samples Using a Hydrophilic Interaction Liquid Chromatography–Inductively Coupled Plasma High-Resolution Mass Spectrometry (HILIC-ICP-HRMS) Method

Here, in this study, a novel method using hydrophilic interaction liquid chromatography (HILIC) coupled with inductively coupled plasma high-resolution mass spectrometry (ICP-HRMS) was introduced for the quantification of phospholipids in oil samples. The method employed a bridged ethyl hybrid (BEH) stationary phase HILIC column with a tetrahydrofuran (THF)/water mobile phase, enhancing the solubility and detection of phospholipids. During the study, a gradient/matrix effect on ICP-HRMS sensitivity was observed and successfully compensated for experimentally, ensuring reliable quantification results. This approach has proven effective for a wide range of different oil samples including vegetable oils, animal fats, and phospholipid supplements. Notably, this method allowed the direct quantification of phospholipids in oil samples, bypassing the need for prior sample preparation methods, such as solid phase extraction (SPE), thereby streamlining the analytical process. The precision, accuracy, and reduced need for extensive sample preparation offered by this method mark a significant advancement in lipids analysis. Its robustness and broad applicability have substantial implications for industries such as food and renewable energy production, where both efficient and accurate lipid identification and quantification are crucial.

09 BIOMASS FUELS↗

Zinc isotope constraints on the cycling of carbon in the Bermuda mantle source

Volatile recycling and storage in the mantle transition zone (MTZ) is important for the refertilization of the upper mantle and is associated with the generation of high-µ (HIMU, where µ is 238 U/ 204 Pb) mantle. One way to probe the MTZ and the processes associated with mantle convection is to sample lavas that originate from the shallow mantle and were contaminated by upwelling from the MTZ, such as at the previously proposed shallow plume of Bermuda. Here we present the first δ 66 Zn isotopic compositions of Bermuda silica-undersaturated and silica-saturated lavas to explore the origin of the carbon-rich lithologies and the genesis of the large seamount found in the western North Atlantic Ocean. Contrasting with global δ 66 Zn data sets, our results (δ 66 Zn between 0.24 ± 0.04 and 0.41 ± 0.04) do not support direct sampling of recycled marine carbonates in the Bermuda HIMU mantle. Instead, we show that δ 66 Zn fractionation toward higher values is associated with magmatic processes and incorporation of carbon sourced from deep fluids associated with the formation of carbonatites. These carbon-rich fluids are likely sourced from the metasomatic reactions between the subducted cold slab of the Iapetus oceanic lithosphere ca. 500 Ma and the thickened continental lithospheric mantle of Pangea. Melting of this metasomatized mantle was triggered by the arrival of the Farallon slab to the eastern North American margin in the late Cenozoic via shallow convection.

Mazza, Sarah E. [Smith College, Northampton, MA (U↗

A Perspective on the Milky Way Bulge Bar as Seen from the Neutron-capture Elements Cerium and Neodymium with APOGEE

Abstract This study probes the chemical abundances of the neutron-capture elements cerium and neodymium in the inner Milky Way from an analysis of a sample of ∼2000 stars in the Galactic bulge bar spatially contained within ∣X Gal ∣ < 5 kpc, ∣Y Gal ∣ < 3.5 kpc, and ∣Z Gal ∣ < 1 kpc, and spanning metallicities between −2.0 ≲ [Fe/H] ≲ +0.5. We classify the sample stars into low- or high-[Mg/Fe] populations and find that, in general, values of [Ce/Fe] and [Nd/Fe] increase as the metallicity decreases for the low- and high-[Mg/Fe] populations. Ce abundances show a more complex variation across the metallicity range of our bulge-bar sample when compared to Nd, with ther-process dominating the production of neutron-capture elements in the high-[Mg/Fe] population ([Ce/Nd] < 0.0). We find a spatial chemical dependence of Ce and Nd abundances for our sample of bulge-bar stars, with low- and high-[Mg/Fe] populations displaying a distinct abundance distribution. In the region close to the center of the MW, the low-[Mg/Fe] population is dominated by stars with low [Ce/Fe], [Ce/Mg], [Nd/Mg], [Nd/Fe], and [Ce/Nd] ratios. The low [Ce/Nd] ratio indicates a significant contribution in this central region fromr-process yields for the low-[Mg/Fe] population. The chemical pattern of the most metal-poor stars in our sample suggests an early chemical enrichment of the bulge dominated by yields from core-collapse supernovae andr-process astrophysical sites, such as magnetorotational supernovae.

Astronomy & Astrophysics↗

Quasi-static to Dynamic Mechanical Response and Microstructure Development of Tantalum-Tungsten Alloys

Lawrence Livermore National Laboratory (LLNL) is interested the quasi-static to dynamic mechanical response and microstructure evolution of tantalum-tungsten (Ta-W) alloys (Ta-2.5W, Ta-5W, and Ta-10W, wt.%) made by conventional wrought processing and additive manufacturing (AM). This work scope was performed at the Colorado School of Mines (Mines) and included quasi-static (e.g., 10 -3 s -1 ) mechanical testing in tension and compression, along with selected high strain rate (Kolsky) pressure bar testing in compression (e.g., 10 3 s -1 ), with and without temperature variations in some instances. Complementary microstructure characterization was performed on undeformed and deformed samples to understand the role of processing on microstructural evolution and the deformation mechanisms that impact the mechanical response with variations in strain rate, temperature, and strain state (e.g., tension versus compression in selected examples). The wrought material provided by LLNL from Viridis Materials was found to have unrecrystallized regions within the microstructure, which led to unexpected results relative to previously reported properties for Ta-W. The AM Ta-2.5W (wt.%) material provided by LLNL was found to have higher compressive strength than wrought Ta-2.5W (wt.%), which is hypothesized to be due to differences in the crystallographic texture between the two materials. This project partially supported several postdocs and graduate students at Mines.

36 MATERIALS SCIENCE↗

Constraining Cross Section and Beam Systematics for Future NOvA Sterile Neutrino Search

he NOvA (NuMI Off-axis $\nu_e$ Appearance) experiment measures neutrino oscillations in a nearly pure muon (anti)neutrino beam over a 810 km baseline. We search for sterile neutrino driven oscillations in both neutral current (NC) and charged current (CC) muon neutrino samples. A deficit of NC interactions at the Far Detector could occur since sterile neutrinos do not couple to the Z boson. A modulation of the muon neutrino disappearance probability is possible due to a new mass splitting. The NuMI beam line is capable of running in forward or reverse horn current modes, which creates a neutrino or antineutrino beam, respectively. Samples collected in forward horn current (FHC) or reverse horn current (RHC) modes are subject to beam optics uncertainties, in addition to cross section and hadron production uncertainties. We explore the ategories. To tackle reduction of cross section uncertainties, we study how splitting the neutral current (NC) samples can impact the systematic reduction and its implications on oscillation parameters, leveraging results from Monte Carlo simulations. Additionally, we look at the possibility of using a $\nu$-on-e scattering sample. The $\nu$-on-e scattering process has minimal cross section uncertainties compared to that of neutrino-nucleus interactions, which can help constrain the overall flux normalizations. Using both the ZHC sample, NC split sample, and $\nu$-on-e scattering sample jointly with FHC samples would reduce the focusing uncertainties and the hadron production uncertainties in the future sterile analysis.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Prediction of Silicon Content in a Blast Furnace via Machine Learning: A Comprehensive Processing and Modeling Pipeline

Silicon content plays an important role in determining the operational efficiency of blast furnaces (BFs) and their downstream processes in integrated steelmaking; however, existing sampling methods and first-principles models are somewhat limited in their capability and flexibility. Current data-based prediction models primarily rely on a limited set of manually selected furnace parameters. Additionally, different BFs present a diverse set of operating parameters and state variables that are known to directly influence the hot metal’s silicon content, such as fuel injection, blast temperature, and raw material charge composition, among other process variables that have their own impacts. The expansiveness of the parameter set adds complexity to parameter selection and processing. This highlights the need for a comprehensive methodology to integrate and select from all relevant parameters for accurate silicon content prediction. Providing accurate silicon content predictions would enable operators to adjust furnace conditions dynamically, improving safety and reducing economic risk. To address these issues, a two-stage approach is proposed. First, a generalized data processing scheme is proposed to accommodate diverse furnace parameters. Second, a robust modeling pipeline is used to establish a machine learning (ML) model capable of predicting hot metal silicon content with reasonable accuracy. The method employed herein predicted the average Si content of the upcoming furnace cast with an accuracy of 91% among 200 target predictions for a specific furnace provisioned by the XGBoost model. This prediction is achieved using only the past shift’s operating conditions, which should be available in real time. This performance provides a strong baseline for the modeling approach with potential for further improvement through provision of real-time features.

Chemistry↗

Optimising the processing and storage of visibilities using lossy compression

The next-generation radio astronomy instruments are providing a massive increase in sensitivity and coverage, largely through increasing the number of stations in the array and the frequency span sampled. The two primary problems encountered when processing the resultant avalanche of data are the need for abundant storage and the constraints imposed by I/O, as I/O bandwidths drop significantly on cold storage. An example of this is the data deluge expected from the SKA Telescopes of more than 60 PB per day, all to be stored on the buffer filesystem. While compressing the data is an obvious solution, the impacts on the final data products are hard to predict. In this paper, we chose an error-controlled compressor – MGARD – and applied it to simulated SKA-Mid and real pathfinder visibility data, in noise-free and noise-dominated regimes. As the data have an implicit error level in the system temperature, using an error bound in compression provides a natural metric for compression. MGARD ensures the compression incurred errors adhere to the user-prescribed tolerance. To measure the degradation of images reconstructed using the lossy compressed data, we proposed a list of diagnostic measures, exploring the trade-off between these error bounds and the corresponding compression ratios, as well as the impact on science quality derived from the lossy compressed data products through a series of experiments. We studied the global and local impacts on the output images for continuum and spectral line examples. We found relative error bounds of as much as 10%, which provide compression ratios of about 20, have a limited impact on the continuum imaging as the increased noise is less than the image RMS, whereas a 1% error bound (compression ratio of 8) introduces an increase in noise of about an order of magnitude less than the image RMS. For extremely sensitive observations and for very precious data, we would recommend a 0.1% error bound with compression ratios of about 4. These have noise impacts two orders of magnitude less than the image RMS levels. At these levels, the limits are due to instabilities in the deconvolution methods. We compared the results to the alternative compression tool DYSCO, in both the impacts on the images and in the relative flexibility. MGARD provides better compression for similar error bounds and has a host of potentially powerful additional features.

Techniques: interferometric↗

Computationally efficient and error aware surrogate construction for numerical solutions of subsurface flow through porous media

Limiting the injection rate to restrict the pressure below a threshold at a critical location can be an important goal of simulations that model the subsurface pressure between injection and extraction wells. The pressure is approximated by the solution of Darcy’s partial differential equation for a given permeability field. The subsurface permeability is modeled as a random field since it is known only up to statistical properties. This induces uncertainty in the computed pressure. Solving the partial differential equation for an ensemble of random permeability simulations enables estimating a probability distribution for the pressure at the critical location. These simulations are computationally expensive, and practitioners often need rapid online guidance for real-time pressure management. An ensemble of numerical partial differential equation solutions is used to construct a Gaussian process regression model that can quickly predict the pressure at the critical location as a function of the extraction rate and permeability realization. The Gaussian process surrogate analyzes the ensemble of numerical pressure solutions at the critical location as noisy observations of the true pressure solution, enabling robust inference using the conditional Gaussian process distribution. Our first novel contribution is to identify a sampling methodology for the random environment and matching kernel technology for which fitting the Gaussian process regression model scales as O ( n log n ) instead of the typical O ( n 3 ) rate in the number of samples n used to fit the surrogate. The surrogate model allows almost instantaneous predictions for the pressure at the critical location as a function of the extraction rate and permeability realization. Our second contribution is a novel algorithm to calibrate the uncertainty in the surrogate model to the discrepancy between the true pressure solution of Darcy’s equation and the numerical solution. Finally, although our method is derived for building a surrogate for the solution of Darcy’s equation with a random permeability field, the framework broadly applies to solutions of other partial differential equations with random coefficients.

54 ENVIRONMENTAL SCIENCES↗

Preparation of a 73 As source sample for application in an offline ion source

For the generation of beams with the offline ion source at the Facility for Rare Isotope Beams (FRIB), suitable source samples are required. Arsenic-73 is a frequently requested user beam due to its significance in nuclear structure studies and astrophysics. In this work, we outline the process of preparing a 73 As source sample, containing (5.76 ± 0.37)∗10 14 atoms of 73 As, which was successfully used to generate a 73 As beam for a multi-day user experiment. Silver arsenate was chosen as the chemical form, due to its favorable volatility within the designated operating temperature range. We refined the precipitation method using stable arsenic prior to its application with the 73 As sample, resulting in precipitation yields of (99.4 ± 4.5)%.

As-73↗

Rapid measurement of soluble xylo-oligomers using near-infrared spectroscopy (NIRS) and multivariate statistics: calibration model development and practical approaches to model optimization

Rapid monitoring of biomass conversion processes using techniques such as near-infrared (NIR) spectroscopy can be substantially quicker and less labor-, resource-, and energy-intensive than conventional measurement techniques such as gas or liquid chromatography (GC or LC) due to the lack of solvents and preparation methods, as well as removing the need to transfer samples to an external lab for analytical evaluation. The purpose of this study was to determine the feasibility of rapid monitoring of a biomass conversion process using NIR spectroscopy combined with multivariate statistical modeling, and to examine the impact of (1) subsetting the samples in the original dataset by process location and (2) reducing the spectral range used in the calibration model on model performance. We develop multivariate calibration models for the concentrations of soluble xylo-oligosaccharides (XOS), monomeric xylose, and total solids at multiple points in a biomass conversion process which produces and then purifies XOS compounds from sugar cane bagasse. A single model using samples from multiple locations in the process stream showed acceptable performance as measured by standard statistical measures. However, compared to the single model, we show that separate models built by segregating the calibration samples according to process location show improved performance. We also show that combining an understanding of the sample spectra with simple multivariate analysis tools can result in a calibration model with a substantially smaller spectral range that provides essentially equal performance to the full-range model. We demonstrate that real-time monitoring of soluble xylo-oligosaccharides (XOS), monomeric xylose, and total solids concentration at multiple points in a process stream using NIR spectroscopy coupled with multivariate statistics is feasible. Segregation of sample populations by process location improves model performance. Models using a reduced spectral range containing the most relevant spectral signatures show very similar performance to the full-range model, reinforcing the importance of performing robust exploratory data analysis before beginning multivariate modeling.

09 BIOMASS FUELS↗

Data and scripts associated with “Non-random processes impacting organic matter chemistry are maximized in mid-order streams”

NOTE: The manuscript associated with this data package is currently in review. The data may be revised based on reviewer feedback. Upon manuscript acceptance, this data package will be updated with the final dataset and additional metadata. This data package is associated with the publication “Non-random processes impacting organic matter chemistry are maximized in mid-order streams” submitted to Limnology and Oceanography (L&O) by Danczak et al. (in review). This package contains data and scripts used to investigate dissolved organic matter (DOM) molecular chemistry and diversification processes across 47 surface-water sampling sites in the Yakima River Basin, Washington, USA, during an August 2021 sampling campaign. The package contains analyses of ultrahigh-resolution Fourier transform ion cyclotron resonance mass spectrometry (FTICR-MS), geochemical measurements, geospatial attributes, molecular diversity, and meta-metabolome ecological null models needed to reproduce the main manuscript results. The underlying field data were pulled from exising data packages at https://doi.org/10.15485/1892052 (Fulton et al., 2022) and https://doi.org/10.15485/1898914 (Grieger et al., 2022). For details on how to navigate data packages generated by this project, see https://data.ess-dive.lbl.gov/portals/PNNLRiverCorridorSFA/About. We thank the following organizations for providing access to field locations for sample collection: the United States Forest Service, Washington Department of Fish and Wildlife, Washington Department of Natural Resources, the Confederated Tribes and Bands of the Yakama Nation, and the Cowiche Canyon Conservatory. Research was conducted under Washington State Parks and Recreation Commission Scientific Research Permit #210901. We are grateful to the Yakama Nation Tribal Council and Yakama Nation Fisheries for their collaboration in facilitating sample collection and ensuring data usage aligns with their values and worldview. This data package contains an R-Markdown file for analyses and five folders: (1) Data, (2) Geospatial Data, (3) Supplemental_Files, (5) Figures_pdf, (4) and src. The Data folder contains tabular inputs and derived files used in the manuscript analysis. The Geospatial Data folder contains climate and water-balance, hydrologic, land-cover, population/regional water-use, stream, topographic, and stream-order attribute CSV files. The src folder contains scripts used to process data, run analyses, and generate figures. The Figures_pdf folder contains manuscript figure outputs. The Supplemental_Files folder contains supplemental analysis products. All files are .csv, .pdf, .html, .png, .R, .Rmd, .svg, or .tre. This data package is associated with the rcfsa-RC2-SPS_Null_Modeling repository found at https://github.com/river-corridors-sfa/rcfsa-RC2-SPS_Null_Modeling.

54 ENVIRONMENTAL SCIENCES↗