Search NASA⌕ Search

SEARCH · Search NASA

Results for “SAMPLED DATA SYSTEM”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 253 records · Page 14

Measurements of the z > 5 Lyman-α forest flux autocorrelation functions from the extended XQR-30 data set

We present the first observational measurements of the Lyman-α (Ly α) forest flux autocorrelation functions in ten redshift bins from 5.1 ≤ z ≤ 6.0. We use a sample of 35 quasar sightlines at z > 5.7 from the extended XQR-30 data set; these data have signal-to-noise ratios of >20 per spectral pixel. We carefully account for systematic errors in continuum reconstruction, instrumentation, and contamination by damped Ly α systems. With these measurements, we introduce software tools to generate autocorrelation function measurements from any simulation. Our measurements of the smallest bin of the autocorrelation function increase with redshift when normalizing by the mean flux, $\langle{F}\rangle$. This increase may come from decreasing $\langle{F}\rangle$ or increasing mean free path of hydrogen-ionizing photons, λmfp. Recent work has shown that the autocorrelation function from simulations at z > 5 is sensitive to λmfp, a quantity that contains vital information on the ending of reionization. For an initial comparison, we show our autocorrelation measurements with simulation models for recently measured λmfp values and find good agreements. Further work in modelling and understanding the covariance matrices of the data is necessary to get robust measurements of λmfp from this data.

79 ASTRONOMY AND ASTROPHYSICS↗

Measurements of short-lived fission product yields from photofission of 238 U using 13.0 MeV monoenergetic photons

Photon-induced fission product yield (FPY) measurements were conducted on the isotope 238U. Fission was induced using Eγ = 13.0 MeV monoenergetic photons produced by the Triangle Universities Nuclear Laboratory’s (TUNL’s) High Intensity γ-ray Source (HIγS) facility. Short-lived FPYs were measured by performing cyclic activation of the sample using a rapid target transfer system. Following activation, the 238U target was rapidly (0.4 s) transferred to a counting station consisting of two well-shielded high-purity germanium (HPGe) detectors. The irradiation-counting cycle was repeated until the summed data had sufficient statistical accuracy. Twenty-eight unique fission products with half-lives ranging from 1 s to 450 s were identified, and their cumulative FPYs determined. Furthermore, the results are compared with previous independent FPY measurements using inverse kinematics. Good agreement between the data sets is found despite the different excitation energy distributions of the fissioning nucleus in the experiments.

Physics - Nuclear physics and radiation physics↗

Evaluation of an event-driven 3FI ASIC for spectroscopic X-ray detection with synchrotron radiation

The novel design and evaluation on the NSLS-II beamline of the 3FI application-specific integrated circuit (ASIC) bump-bonded to a simple, planar, 2D segmented silicon sensor are presented. The ASIC was developed for full-field fluorescence spectral X-ray imaging (3FI). It is a small-scale prototype that features a square array of 32 × 32 pixels, and the size of the pixels is 100 µm × 100 µm. The ASIC was implemented in a 65 nm CMOS integrated circuit fabrication process. Each pixel incorporates a charge-sensitive amplifier, a shaping filter, a discriminator, a peak detector and a sample-and-hold circuit, allowing detection of events and storage of signal amplitudes. The system operates in a frameless event-driven readout mode, outputting analog values for threshold-triggered events, allowing high-speed multi-element X-ray fluorescence data acquisition. The 3FI ASIC achieves per-channel spectrometric performance at a power consumption of only 200 µW per pixel, with nearly all dissipation confined to the analog front-end. An energy resolution is measured at the level of 308 eV full width at half-maximum (FWHM) at 8.04 keV (Cu Kα), and 138 eV FWHM at 3.69 keV (Ca Kα). This per-pixel capability makes the prototype suitable for in situ trace element microanalysis in biological and environmental studies. Moreover, the frameless architecture of the detector is designed to address limitations of conventional X-ray fluorescence microscopy, which typically requires mechanical scanning, by enabling continuous high-throughput data acquisition in future full-field implementations.

47 OTHER INSTRUMENTATION↗

2010-2012 California Household Travel Survey

The 2010-2012 California Household Travel Survey (CHTS) was administered by the California Department of Transportation, which collected demographic and travel behavior characteristics for residents across the entire state. At the time, it was the largest such regional or statewide survey ever conducted in the United States. Detailed travel behavior information was obtained from more than 42,500 households via multiple data-collection methods, including computer-assisted telephone interviewing, online and mail surveys, wearable (7,574 participants) and in-vehicle (2,910 vehicles) global positioning system devices, and on-board diagnostic sensors that gathered data directly from a vehicle's engine. Details of personal travel behavior were gathered within the region of residence, inter-regionally within the state, and in adjoining states and Mexico. The survey sampling plan was designed to ensure an accurate representation of the entire population of the state. The CHTS included additional features, such as vehicle-acquisition decisions, parking choices, work schedules and flexibility, use of toll lanes/priced facilities, and walk and bicycle trips, to support advanced model development.

1Hz data↗

Evaluation of the Radioactive Material Released in the Harborview Research and Training Building and Some Implications for Emergency Response

On 2 May 2019, during the 137 Cs source recovery operation, a source capsule in a research irradiator containing approximately 77.1 TBq was breached. Based on a geometric reconstruction analysis of the damage to the capsule, approximately 46.3 GBq (0.04%) was impacted by the chop saw (grinder) inside a mobile hot cell on the loading dock at the University of Washington Harborview Research and Training (HRT) Building. A very small fraction of the material impacted, less than 1%, was released from the mobile hot cell and then to the rest of the HRT Building. The objectives of this project were to assess the accidental release of 137 CsCl and its implications related to emergency response methods and the ramifications of 137 CsCl transport. The phenomenology of this event was also compared with past alkali halide dispersal events. The vast number of measurements and samples collected by the remediation contractors, the Department of Energy’s Nuclear Emergency Support Team, and the small number of retrospective samples collected by the authors informed the analysis. The techniques included (1) autoradiography and electron microscopy of samples collected from the HRT Building and the irradiator, (2) 3D visualization of deposition on surfaces and within the ventilation system, and (3) a study of the damage to the source capsule to evaluate the Cs particle size and particle composition due to the grinding accident. Subsequently, the cesium contaminant transport through the numerous pathways in the building was reconstructed to assess the deposition on surfaces as a function of particle size. Furthermore, the implications for emergency response are relevant to data quality and management. A Data Quality Objective guides data collection methods so that they have appropriate accuracy and precision for the intended application. Recommendations were made with respect to the sample collection protocols and archiving of samples.

46 - INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AN↗

Data for Kim et al., "Variations in the optical and molecular composition of dissolved organic matter exported from coastal wetlands"

Knowledge about sources and composition of marsh-derived dissolved organic matter (DOM) is critical for understanding the role of marshes in coastal biogeochemical cycling and the fate of marsh-derived DOM in the ocean. To investigate tidal variability in composition of marsh-derived DOM, Kim et al. examined the optical and molecular characteristics of hourly surface water samples at three tidal creeks in the Chesapeake Bay. Groundwater samples along the terrestrial landscape gradient as well as estuarine water from the adjacent estuary at each site were also collected to help resolve sources of surface water DOM. Samples were collected in summer 2024 at three sites – SWH: Sweet Hall Marsh, GCW: Kirkpatrick Marsh, and GWI: Goodwin Islands – which are part of synoptic sites in the Chesapeake Bay region of the COMPASS-FME (Coastal Observations, Mechanisms, and Predictions Across Systems and Scales - Field, Measurements, and Experiments) project. Surface water samples were collected hourly over a 48-hour period at each site. Groundwater and estuarine water samples were collected once. This dataset includes- Surface water depth and salinity- Dissolved organic carbon (DOC) and total dissolved nitrogen (TDN) concentrations- Optical indices and relative composition of parallel factor analysis (PARAFAC) components- High resolution mass spectrometry data.

54 ENVIRONMENTAL SCIENCES↗

Effects of 9.5 years warming on SOC concentration and composition in bulk soil and density fractions

Original data of whole-soil warming experiment after 9.5 years at Blodgett Forest Research Station. The Blodgett Forest is a mixed coniferous temperate forest with Mediterranean climate. The annual air temperature is 12.5℃ and the annual precipitation is 1774 mm yr-1- The soil is mesic ultic Alfisol of granitic origin, equivalent to Dystric Cambisol according to The World Reference Base for Soil Resources (WRB) system. The soil is warmed down to 1 m at + 4℃ by vertically installed heating cables. At the time of soil sampling on 1 May 2023, the whole-soil warming experiment had been running for approximately 9.5 years, from January 2014 to May 2023. The dataset includes: - Bulk_EA: C, N content, δ13C, and CN ratio of bulk soil; - Density_fractionation: organic carbon concentration, δ13C, and C/N ratio of free light fraction (fLF), occluded ligh fraction (oLF), and heavy fraction (HF); - PCA_DRIFT_AUC: original data of area under the curve (AUC) values of eight carbon bond types integrated on diffuse reflectance infrared fourier transform spectroscopy for each soil sample and soil fraction, which are consequently used for principal component analysis (PCA); - DRIFTS_stability_index: the calculation of aliphatic C–H (3000–2800 cm-1) to aromatic C=C (1670–1600 cm-1) ratios for each bulk soil sample and soil fraction. All data are provided in CSV format and can be viewed using Microsoft Excel.

Climate change↗

User Guide for Sample Reduction at GP-SANS

This manual is intended as a quick guide for data reduction of GP-SANS data. It includes all necessary steps to do the data reduction based on absolute calibration using the open beam method and how to transfer the reduced data to the personal computer system. If any errors are coming up so that the reduction script is not functioning as intended, please contact the instrument scientist.

97 MATHEMATICS AND COMPUTING↗

Bioelectrochemical crossbar architecture screening platform for extracellular electron transfer

Electroactive microbes can serve as living components in bioelectronic devices, where their unique ability to transfer electrons enables applications in sensing, energy conversion, and synthesis, but they remain challenging to engineer because the bioelectrochemical systems (BESs) used for characterization are low throughput. Here, we present a bioelectrochemical crossbar architecture screening platform (BiCASP) that uses stacked and orthogonally arrayed electrodes to enable individual sample selection for characterization in arrayed formats. This device reports on the current generated by electroactive bacteria on the minute timescale, decreasing the time for data acquisition by several orders of magnitude compared to conventional BESs. This device increases the throughput of screening engineered biological components in cells, identifying mutants of the membrane protein wire MtrA in Shewanella oneidensis that retain the ability to support extracellular electron transfer (EET). BiCASP may be integrated with bioelectronics that need directed evolution of electroactive proteins.

Shewanella↗

Performance Evaluation of Vertical Federated Machine Learning Against Adversarial Threats on Wide-Area Control System: Preprint

Federated machine learning (FL) is gaining significant popularity to develop cybersecurity solutions in power grids because of its advanced capability to support decentralized data handing at local devices, its privacy preservation, and its low-bandwidth requirement. However, the evolving adversarial machine learning (AML) threats raise significant concerns for the cybersecurity of FL architectures. The FL-based split neural network (SplitNN) achieves high performance through the decentralized training of local neural network models while preserving data privacy across multiple entities. In this paper, we propose a methodology for evaluating the performance of a vertical FLbased anomaly detector against different types of AML attacks, including denial-of-service attacks, adversarial data injection attacks, and replay attacks on the trained local models deployed in the grid network. For a case study, we consider the modified IEEE 13-bus system, and we develop SplitNN-based binary and multiclass classification models to detect, locate, and identify different types of data integrity attacks on the volt-watt control with two pooling layers: maximum pooling and AvgPool. Our experimental results, computed through performance metrics, reveal that the severity of these AML attacks varies with the integrated pooling mechanism, the type of classification model, and the nature of the cyberattack. Further, the AML attacks negatively impacted the prediction time per sample for the pretrained SplitNN during the online testing.

adversarial threats↗

ARM Trajectories Data Set Value-Added Product Report

The U.S. Department of Energy Atmospheric Radiation Measurement (ARM) user facility’s ARM Trajectories Data Set (ARMTRAJ) Value-Added Product (VAP) provides trajectory data sets initialized at ARM deployment coordinates and configured using ARM data sets. The six trajectory data sets support aerosol, cloud, planetary boundary layer, and related research (aerosol-cloud interactions, etc.), as well as studies using ARM Aerial Facility (AAF) and tethered balloon system (TBS) measurements. Trajectory calculations use the Hybrid Single-Particle Lagrangian Integrated Trajectory (HYSPLIT) model informed by the European Centre for Medium-Range Weather Forecasts (ECMWF) fifth-generation atmospheric reanalysis (ERA5) data set at its highest spatial resolution (~31 km). HYSPLIT also runs at multiple initial starting locations surrounding ARM deployments (in latitude/longitude and/or vertical coordinates), facilitating an ensemble for each sample in the data sets. The ensemble mean and variability reported in ARMTRAJ improve the fidelity and provide uncertainty estimates of trajectory coordinates, thermodynamic properties, and other output fields.

54 ENVIRONMENTAL SCIENCES↗

Molecular Dynamics Simulation of Complex Reactivity with the Rapid Approach for Proton Transport and Other Reactions (RAPTOR) Software Package

Simulating chemically reactive phenomena such as proton transport on nanosecond to microsecond and beyond time scales is a challenging task. Ab initio methods are unable to currently access these time scales routinely, and traditional molecular dynamics methods feature fixed bonding arrangements that cannot account for changes in the system’s bonding topology. The Multiscale Reactive Molecular Dynamics (MS-RMD) method, as implemented in the Rapid Approach for Proton Transport and Other Reactions (RAPTOR) software package for the LAMMPS molecular dynamics code, offers a method to routinely sample longer time scale reactive simulation data with statistical precision. RAPTOR may also be interfaced with enhanced sampling methods to drive simulations toward the analysis of reactive rare events, and a number of collective variables (CVs) have been developed to facilitate this. Key advances to this methodology, including GPU acceleration efforts and novel CVs to model water wire formation are reviewed, along with recent applications of the method which demonstrate its versatility and robustness.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

From Reads to Function Workshop - Milano 2026

The Bicocca Sampling Days (BSDs) model offers a reproducible “citizen science” framework integrating research, education, and public engagement through large-scale microbiome sampling, followed by a workshop of data analysis on select samples. We identified 9 bacterial and archaeal metagenome-assembled genomes from six soil samples across three separate sampling days in two approaches with indidivual sample and replicate co-assembly spanning three unique classes, providing genomic insights into microbial nutrient cycling in these systems.

59 BASIC BIOLOGICAL SCIENCES↗

Mining Thermophile Photosynthesis Genes: A Synthetic Operon Expressing Chloroflexota Species Reaction Center Genes in Rhodobacter sphaeroides

Photosynthesis is the foundation of the vast majority of life systems, and is therefore the most important bioenergetic process on earth. The greatest diversity of photosynthetic systems is found in microorganisms. However, our understanding of the biophysical and biochemical processes that transduce light into chemical energy is derived from a relatively small subset of proteins from microbes that are amenable to cultivation, in contrast to the huge number of predicted proteins that catalyze the initial photochemical reactions deposited in databases, such as from metagenomics. We describe the use of a Rhodobacter sphaeroides laboratory strain for the expression of heterologous photosynthesis genes to demonstrate the feasibility of mining this resource, focusing on hot spring Chloroflexota gene sequences. Using a synthetic operon of genes, we produced a photochemically active complex of reaction center proteins in our biological system. We also present bioinformatic analyses of anoxygenic type II reaction center sequences from metagenomic samples collected from hot (42–90 °C) springs available through the JGI IMG database, to generate a resource of diverse sequences that are potentially adapted to photosynthesis at such temperatures. These data provide a view into the natural diversity of anoxygenic photosynthesis, through a lens focused on high-temperature environments. The approach we took to express such genes can be applied for potential biotechnology purposes as well as for studies of fundamental catalytic properties of these heretofore inaccessible protein complexes.

Chloroflexota↗

MULTI-LEADER: MULTI-source LEarning-Accelerated Design of high-Efficiency multi-stage compRessor (Final Technical Report)

The objective of MULTI-LEADER is to cut design costs by 80% while generating more energy-efficient designs of multi-stage compressors by developing and implementing novel machine learning (ML) techniques, which enable faster and fewer design iterations, improved solver performance, and concurrent multi-disciplinary design. Current industrial practices for the design of multi-stage compressors involve simulation-based design optimization with successive levels of model fidelity, iteratively evaluated between distinct disciplines, one stage at a time to tackle the high dimensional design variations. This project addresses these key design challenges: (1) concurrent optimization of multiple stages under many non-linear constraints; (2) multitude of evaluation of high-fidelity and expensive solvers and their gradients during optimization convergence in high-dimensional design; (3) multi-disciplinary design to maximize aerodynamic performance while guaranteeing structural integrity and additive manufacturability; (4) utilization of multiple fidelity of solvers with disparate parameterization and modeling assumptions. MULTI-LEADER achieved more than 5x speed up in detailed design of more energy-efficient compressors via these machine learning (ML) innovations: (i) rapid design surrogates by multi-source learning from diverse fidelities across multiple disciplines, (ii) physics-constrained data-augmented modeling for improved empiricism, (iii) generative manifold embedding for high dimensional concurrent design without gradient information; (iv) budget-constrained fidelity-adaptive sampling towards fewer design iterations.

33 ADVANCED PROPULSION SYSTEMS↗

Addressing GPU memory limitations for Graph Neural Networks in High-Energy Physics applications

Introduction Reconstructing low-level particle tracks in neutrino physics can address some of the most fundamental questions about the universe. However, processing petabytes of raw data using deep learning techniques poses a challenging problem in the field of High Energy Physics (HEP). In the Exa.TrkX Project, an illustrative HEP application, preprocessed simulation data is fed into a state-of-art Graph Neural Network (GNN) model, accelerated by GPUs. However, limited GPU memory often leads to Out-of-Memory (OOM) exceptions during training, due to the large size of models and datasets. This problem is exacerbated when deploying models on High-Performance Computing (HPC) systems designed for large-scale applications. Methods We observe a high workload imbalance issue during GNN model training caused by the irregular sizes of input graph samples in HEP datasets, contributing to OOM exceptions. We aim to scale GNNs on HPC systems, by prioritizing workload balance in graph inputs while maintaining model accuracy. Our paper introduces diverse balancing strategies aimed at decreasing the maximum GPU memory footprint and avoiding the OOM exception, across various datasets. Results Our experiments showcase memory reduction of up to 32.14% compared to the baseline. We also demonstrate the proposed strategies can avoid OOM in application. Additionally, we create a distributed multi-GPU implementation using these samplers to demonstrate the scalability of these techniques on the HEP dataset. Discussion By assessing the performance of these strategies as data loading samplers across multiple datasets, we can gauge their effectiveness in both single-GPU and distributed environments. Our experiments, conducted on datasets of varying sizes and across multiple GPUs, broaden the applicability of our work to various GNN applications that handle input datasets with irregular graph sizes.

Lee, Claire Songhyun↗

Dataset 1: A National and City Dataset on Human Factors in Pooled Rideshare, 2021

Dataset 1: A National and City Dataset on Human Factors in Pooled Rideshare, 2021. Dataset Description: Pooled Rideshare Acceptance Survey - Phase 1 (2021, N = 5,385). This dataset captures responses from a nationally representative sample of 5,385 adults across the United States to understand public acceptance, preferences, and behavioral intentions related to pooled rideshare (PR) services. The primary objective of this research is to provide actionable insights to inform the design, deployment, and policy development of sustainable shared mobility systems. Data was collected via an online survey administered through a national panel provider. Participants ranged in age from 18 to 95 years, and representation from all U.S. regions. The survey instrument was designed to explore numerous dimensions related to PR adoption including demographic traits, current travel habits, rideshare familiarity, trust, safety, environmental attitudes, and user experience preferences. Both rideshare users and non-users were included, offering a diverse range of perspectives. - Phase_1_Final - The dataset includes survey items developed from literature reviews, and prior field studies. Each row represents an individual respondent, and each column corresponds to a variable such as willingness to use pooled rideshare, attitudes toward specific service features, and sociodemographic data. The data is available in both .CSV and .SAV formats. - Phase_1_Final_MapFile - The accompanying data dictionary explains all variable labels, response scales, and codes. An .XLSX format of the full survey instrument is also included to support interpretation and reuse of the dataset.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Development and implementation of high-throughput proteomic and metabolomics assays by using advanced chromatographic and mass spectrometric systems (CRADA Final Report)

The mission of this CRADA with Agilent was to couple powerful MS platforms (QQQ, IM-QTOFMS) with Agilent’s novel Ultra-High-Performance Liquid Chromatography (UHPLC) fast metabolomic workflows and perform ABF Machine Learning (ML) to generated datasets. Agilent transferred UHPLC methods to PNNL and LBNL and methods were implemented and demonstrated in both labs, achieving total acquisition times of < 10 min. Metabolites analyzed using Agilent’s shared methods included metabolites from central carbon metabolism, common across hosts, and metabolites unique to engineered strains. Standards were acquired in an UHPLC-Drift Tube Ion Mobility Mass Spectrometer (DTIMS) system for the first time within the context of ABF and methods were optimized based on Agilent’s protocols. Samples from ABF hosts Pseudomonas putida, Aspergillus pseudoterreus, Aspergillus niger and Rhodosporidium toruloides were analyzed using the UHPLC-DTIMS platform for a total of 276 runs. A data analysis workflow compatible with the Experimental Data Depot (EDD) and completely shareable was developed for the acquired UHPLC-DTIMS data. Samples were analyzed using a Data Independent Acquisition Approach (DIA), which for most of the standards provided more transitions therefore increasing detection confidence. Using the data acquired by PNNL, LBNL, and Agilent’s specifications from previous ML projects, SNL applied an ensemble ML strategy to pick the best performing model for automated LC-method selection. Finally, with the contribution of the participant labs and Agilent, SNL developed an Automated Method Selection (AMS) software tool to predict the best liquid chromatography method for analysis of any new molecules of interest. Samples with novel pathways and new metabolite targets of interest are generated at a high pace in the ABF. Overall, the project advanced rapid metabolomics by combining liquid chromatography, ion mobility spectrometry, and data-independent mass spectrometry with machine learning. This multidimensional approach uses retention time, collision cross-section, precursor mass, and fragment-ion information to distinguish chemically similar metabolites that can be difficult to resolve using conventional liquid- or gas-chromatography methods. The resulting workflow also provided automated metabolite-identification error estimates, addressing a recognized need for statistical confidence measures in metabolomics.

Petzold, Christopher [Lawrence Berkeley National L↗