Search NASA⌕ Search

SEARCH · Search NASA

Results for “data availability”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 415 records · Page 23

Data and scripts associated with the manuscript "Organic Molecules are Deterministically Assembled in River Sediments"

This data package is associated with the publication "Organic Molecules are Deterministically Assembled in River Sediments" submitted to Scientific Reports (Stegen et al., 2024). The study applies community ecology methods to dissolved organic matter (DOM) chemistry from variably inundated riverbed sediments to uncover principles governing DOM composition at a reach-scale. This data package documents the workflow used to process and generate the main findings in the manuscript. The R scripts reference the raw, unprocessed Fourier transform ion cyclotron resonance mass spectrometry (FTICR-MS) data from another data package, available on ESS-DIVE at https://data.ess-dive.lbl.gov/view/doi:10.15485/1834208. The scripts then process the raw FTICR-MS data and generate the findings and figures presented in the associated manuscript. In brief, this study demonstrates that DOM assemblages in variably inundated sediments are primarily governed by deterministic variable selection, including sediment moisture effecting the degree of deterministic assembly. See the manuscript for more details pertaining to interpretation and implications of the findings. This data package is associated with the GitHub repository found at https://github.com/WHONDRS-Hub/ECA_2020_Sed.This data package is comprised of 6 scripts and 7 folders. The file-level metadata file (file ending in "flmd.csv") lists all files contained in this data package and descriptions for each. The data dictionary (file ending in "dd.csv) describes all tabular data columns and their respective definitions and units. The FTICR_Processing_Scripts produce the outputs found in the "Processed_Data" folder. The remaining scripts (located in the parent directory) produce the outputs found in the following four folders: (1) "MCD_Dendrograms", "MCD_Randomizations", "MCD_bNTI_Outcomes", and "OM_Null_Modeling". The fifth script additionally takes the three comma-separated values (CSV) files found in the parent directory as input ("VGC_texture.csv", "merged_weights.csv", and "ECA2_FTICR_BetaDisp.csv"). The outputs of each of the five scripts serve as the input to the following script, with the final outputs stored in the folder "OM_Null_Modeling".

54 ENVIRONMENTAL SCIENCES↗

WHONDRS 2016 Sediment Organic Matter Characterization Data from Streams across HJ Andrews Experimental Forest, Oregon

This dataset supports a broader synoptic effort to map morphological, hydrological, chemical, and biological conditions across a fifth-order mountain stream network. Samples were generated through a collaborative synoptic sampling effort in 2016. The dataset provides sediment Fourier Transform Ion Cyclotron Resonance Mass Spectrometry (FTICR-MS) from 60 sites across the HJ Andrews Experimental Forest, Oregon (https://andrewsforest.oregonstate.edu). Related data were collected as part of the event and were published separately in collaboration with other team members. The data are available at http://www.hydroshare.org/resource/ea6c0832885a46c3939e7bb22e48e754 and are described within https://doi.org/10.5194/essd-11-1567-2019 (Ward et al., 2019). The hydroshare data package contains processed FTICR-MS data from the samples included in this data package. The data were processed via Formultitude (previously called Formularity; https://github.com/PNNL-Comp-Mass-Spec/Formultitude). However, we have re-processed the data using Core-MS and included it in this data package. Additional related data collected in 2025 from a similar effort can be found at https://data.ess-dive.lbl.gov/datasets/doi:10.15485/3023310 and http://www.hydroshare.org/resource/b274c4a234bf4b12b7cb8a54a696c629. For details on how to navigate data packages generated by this project, see https://data.ess-dive.lbl.gov/portals/PNNLRiverCorridorSFA/About. In addition to a readme, this data package also includes a file-level metadata (FLMD) file that describes each file and a data dictionary (DD) that describes all column/row headers and variable definitions. This dataset is comprised of (1) a folder of sample data; (2) data dictionary; (3) file-level metadata; (4); (5) coordinates; and (6) readme. The sample data subfolder contains 12 Tesla (12T) FTICR-MS data. This folder contains the processed data and three subfolders, one containing the .xml files, one containing the CoreMS output files, and the other containing instructions and scripts for processing the files in CoreMS (https://github.com/EMSL-Computing/CoreMS). All files are .csv, .pdf, .R, .xml, .Rmd, .py, .cal, or .json.

Biogeochemistry↗

Spatial Seal Database for Prospective Storage Resources in the USA

The goal of the Spatial Seal Database for Prospective Storage Resources in the USA is to provide relevant information and spatial extents of caprock and seal rocks. A lack of aggregated information is readily available that focuses on the caprock and seal units within sedimentary basins. The EPA class VI permit requires an assessment of the confining zone as part of submitting a permit. The data catalog of seal unit names and relevant properties with the seal spatial extent database aims to help provided important data for carbon storage based assessments. The data catalog and database are designed to show what seal data is available in a sedimentary basin and guide stakeholders to the original data source for those datasets.

Pantaleone, Scott↗

Unraveling trace anomaly of supradense matter via neutron star compactness scaling

The trace anomaly Δ ≡ 1/3 −𝑃/𝜖 =1/3 −𝜙 quantifies the possibly broken conformal symmetry in supradense matter under pressure 𝑃 at energy density 𝜖. Perturbative QCD (pQCD) predicts a vanishing Δ at extremely high energy or baryon densities when the conformal symmetry is realized but its behavior at intermediate densities reachable in neutron stars (NSs) is still very uncertain. The extraction of Δ from NS observations strongly depends on the employed model for nuclear equation of state (EOS). Using the IPAD-TOV method based on an intrinsic and perturbative analysis of the dimensionless (IPAD) Tolman-Oppenheimer-Volkoff (TOV) equations that are further verified numerically by using 10 5 EOSs generated randomly with a metamodel in a very broad EOS parameter space constrained by terrestrial nuclear experiments and astrophysical observations, here we first show that the compactness 𝜉 ≡ 𝐺⁡𝑀 NS /𝑅⁢𝑐 2 ≡ 𝑀 NS /𝑅 of a NS with mass 𝑀 NS and radius 𝑅 scales very accurately with $\bar{Π}$ c ≡ $Π$ c · (1 +18⁢X/25) ≡ X/(1 +3⁢X 2 +4⁢X) · (1 +18⁢X/25) where X ≡ 𝜙 c = 𝑃 c /𝜖 c is the ratio of pressure over energy density at NS centers. The scaling of NS compactness thus enables one to readily read off the central trace anomaly Δ c = 1/3 −X directly from the observational data of either the mass-radius or red-shift measurements. Finally, we then demonstrate indeed that the available NS data themselves from recent X-ray and gravitational wave observations can determine model insensitively the trace anomaly as a function of energy density in NS cores, providing a stringent test of existing NS models and a clear guidance in a new direction for further understanding the nature and EOS of supradense matter.

nuclear astrophysics↗

Landscape analysis of environmental data sources for linkage with SEER cancer patients database

Abstract One of the challenges associated with understanding environmental impacts on cancer risk and outcomes is estimating potential exposures of individuals diagnosed with cancer to adverse environmental conditions over the life course. Historically, this has been partly due to the lack of reliable measures of cancer patients’ potential environmental exposures before a cancer diagnosis. The emerging sources of cancer-related spatiotemporal environmental data and residential history information, coupled with novel technologies for data extraction and linkage, present an opportunity to integrate these data into the existing cancer surveillance data infrastructure, thereby facilitating more comprehensive assessment of cancer risk and outcomes. In this paper, we performed a landscape analysis of the available environmental data sources that could be linked to historical residential address information of cancer patients’ records collected by the National Cancer Institute’s Surveillance, Epidemiology, and End Results Program. The objective is to enable researchers to use these data to assess potential exposures at the time of cancer initiation through the time of diagnosis and even after diagnosis. The paper addresses the challenges associated with data collection and completeness at various spatial and temporal scales, as well as opportunities and directions for future research.

60 APPLIED LIFE SCIENCES↗

GenAI-Based Digital Twins Aided Data Augmentation Increases Accuracy in Real-Time Cokurtosis-Based Anomaly Detection of Wearable Data

Early detection of potential infectious disease outbreaks is crucial for developing effective interventions. In this study, we introduce advanced anomaly detection methods tailored for health datasets collected from wearables, offering insights at both individual and population levels. Leveraging real-world physiological data from wearables, including heart rate and activity, we developed a framework for the early detection of infection in individuals. Despite the availability of data from recent pandemics, substantial gaps remain in data collection, hindering method development. To bridge this gap, we utilized Wasserstein Generative Adversarial Networks (WGANs) to generate realistic synthetic wearable data, augmenting our dataset for training. Subsequently, we use these augmented datasets to implement a cokurtosis-based technique for anomaly detection in multivariate time-series data. Our approach includes a comprehensive assessment of uncertainties in synthetic data compared to the actual data upon which it was modeled, as well as the uncertainty associated with fine-tuning anomaly detection thresholds in physiological measurements. Through our work, we present an enhanced method for early anomaly detection in multivariate datasets, with promising applications in healthcare and beyond. This framework could revolutionize early detection strategies and significantly impact public health response efforts in future pandemics.

Data-Driven Digital Twins↗

Calculation of ion–ion mutual neutralization rate constants using Landau–Zener theory coupled with trajectory simulations for Ar + –Cl − , Br − , I −

In this computational study, we self-consistently calculate the rate constants of mutual neutralization reactions by incorporating the electron transfer probability, using Landau–Zener state transition theory with inputs derived from ab initio quantum chemistry calculations, into classical trajectory simulations. Electronic structure calculations are done using correlation consistent basis sets with multi-reference configuration interaction to map all the molecular electronic states below the ion-dissociation limit as a function of the distance between the reacting species. Our electronic structure calculations have been significantly improved from our previous work through improved selection of molecular electronic configurations maintaining a fine grid of 1a 0 over a wide range of bond lengths and accurate treatment of spin–orbit couplings. Non-adiabatic coupling matrix elements are calculated with the three-point central difference method near each avoided crossing to estimate the exact crossing point R x and coupling parameter H if , which are inputs to the multi-channel Landau–Zener theory to calculate the electron transition probability. Our approach is applied to estimate the mutual neutralization rate constants for the following ion pairs: Ar + –Cl − , Ar + –Br − , Ar + –I − at ∼133 Pa. Furthermore, our predictions are compared against the experimental data reported. It is seen that the improvement in the electronic structure calculation results in excellent agreement between the simulation results and the available experimental data to within a factor of ∼2 or ∼±50%.

Complete-active space self-consistent field↗

Validation of galvanomagnetic and thermomagnetic transport measurements using Standard Reference Material 3451

In the “method of four coefficients,” electrical resistivity (ρ), Seebeck coefficient (S), Hall coefficient (RH), and Nernst coefficient (Q) of a material are measured and typically fit or modeled with theoretical expressions based on Boltzmann transport theory to glean experimental insights into features of electronic structure and/or charge carrier scattering mechanisms in materials. Although well-defined and readily available reference materials exist for validating measurements of ρ and S, none currently exists for R H or Q. We show that measurements of all four transport coefficients—ρ, S, R H , and Q—can be validated using a single reference sample, namely, the low-temperature Seebeck coefficient Standard Reference Material® (SRM) 3451 (composition Bi 2 Te 3+x ) available from the National Institute for Standards and Technology (NIST) without the need for inter-laboratory sample exchange. Here, R H and Q data for NIST SRM 3451 reported here for the temperature range 80–400 K complement the data already available for ρ and S and will therefore be of interest to researchers desiring to validate new or existing galvanomagnetic and thermomagnetic transport properties measurement systems.

47 OTHER INSTRUMENTATION↗

Estimating Critical Customer Outages Resulting from Extreme Hurricanes

US power outage data has been collected by organizations such as Oak Ridge National Laboratory (ORNL) through Environment for Analysis Geo-Located Energy Infrastructure (EAGLE-I: freely available) and poweroutage.us (commercial data: available to purchase). However, these sources do not provide information specific to outages of critical customers. Critical customers include entities, facilities, and individuals whose continuous access to electricity is essential for public safety, emergency response, disaster recovery, the well-being of vulnerable populations, public safety and order, and public utilities such as natural gas, communications, water and sanitation. Identification and geolocation of critical customers is crucial for understanding and addressing the effects of power outages on essential services and ensuring that necessary measures are taken to maintain their operations during power disruptions. This work is a first step towards estimating the occurrences of critical customer outages and developing a critical customer power outage data repository. This work estimates outage incidents of critical customers through spatiotemporal mapping of power outage data, weather data, building data, and critical infrastructure network data. Our results show that critical customer effects vary across different counties. We provide appropriate mathematical explanations and simplifications to define and systematize the proposed approach.

Bhusal, Narayan [ORNL] (ORCID:0000000222752145)↗

Transformation rate maps of dissolved organic carbon in the contiguous US

Riverine dissolved organic carbon (DOC) plays a vital role in regional and global carbon cycles. However, the processes of DOC conversion from soil organic carbon (SOC) and leaching into rivers are insufficiently understood, inconsistently represented, and poorly parameterized, particularly in land surface and Earth system models. As a first attempt to fill this gap, we propose a generic formula that directly connects SOC concentration with DOC concentration in headwater streams, where a single parameter, the transformation rate from SOC in the soil to DOC leaching flux (P r ), accounts for the overall processes governing SOC conversion to DOC and leaching from soils (along with runoff) into headwater streams. We then derive high-resolution P r maps over the contiguous US (CONUS) using SOC data from two different sources: the Harmonized World Soil Database v1.2 (HWSD) and SoilGrids 2.0. Both maps are developed following the same five major steps: (1) selecting independent catchments where observed riverine DOC data are available with reasonable quality; (2) estimating catchment-average SOC for the independent catchments; (3) estimating the P r values for these catchments based on the generic formula and catchment-average SOC; (4) developing a predictive model of P r with machine learning (ML) techniques and catchment-scale climate, hydrology, geology, and other attributes; and (5) deriving a national map of P r based on the ML model. For evaluation, we compare the DOC concentration derived using the P r map and the observed DOC concentration values at evaluation catchments. The resulting mean absolute scaled error and coefficient of determination are 0.73 and 0.47 for the HWSD-based model and 0.58 and 0.72 for the SoilGrids-based model, respectively, suggesting the effectiveness of the overall methodology. Efforts to constrain uncertainty and evaluate sensitivity of P r to different factors are discussed. To illustrate the use of such maps, we derive a riverine DOC concentration reanalysis dataset over CONUS. The two P r maps, robustly derived and empirically validated, lay a critical cornerstone for better simulating the terrestrial carbon cycle in land surface and Earth system models. Our findings not only set a foundation for improving our predictive understanding of the terrestrial carbon cycle at the regional and global scales, but also hold promises for informing policy decisions related to decarbonization and climate change mitigation. The data presented in this study are publicly available at https://doi.org/10.5281/zenodo.14563816 (Li et al., 2024).

54 ENVIRONMENTAL SCIENCES↗

Deimos: HALEU TRISO Heated Critical Experiment Data

Deimos was the first critical experiment using high-assay low-enriched uranium (HALEU) TRistructural ISOtropic (TRISO) fuel in over 40 years. HALEU TRISO is the desired fuel form for many of the advanced reactor designs in development; however, very little experimental data are available for this fuel type. Deimos was designed to utilize existing HALEU TRISO fuel in a large graphite moderator to obtain nuclear and reactor physics data to fill the gaps surrounding this fuel type and enrichment. In addition to cold critical data, three separate heated experiments were conducted to measure the temperature reactivity coefficient for this type of system. These measured coefficients were then compared to simulated coefficients to a first level order of fidelity. This comparison showed very good agreement for the experiment where only the inner core was heated and good agreement for the other two configurations, which included heating portions of the outer core. Less agreement when the outer core was heated is attributed to potential heating in the beryllium reflector, which has a positive temperature reactivity coefficient and was unaccounted for in the first-order models. Future heated experiments with Deimos will include temperature monitoring of the beryllium reflector to account for beryllium heating in the simulations.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Spurious solar-wind effects on acceleration noise in LISA Pathfinder

Spurious solar-wind effects are a potential noise source in future Laser Interferometer Space Antenna (LISA) measurements. One noise coupling mechanism is constrained by estimating solar-wind effects on acceleration noise in LISA Pathfinder (LPF). While LISA is designed for drag-free differential measurement, predicting the realistic impact both bounds the operational environment and assesses whether LISA could provide serendipitous space-weather observations. Data from NASA's Advanced Composition Explorer (ACE), situated at the L1 Lagrange point, serves as a reliable source of solar-wind data. The data sets are compared over the 114 d time period from 1 March 2016 to 23 June 2016. This period gives the longest readily-available open data set, without interference from other commissioning activities. To evaluate space weather effects, the data from both satellites are formatted, gap-filled/interpolated, and fast-Fourier transformed for amplitude spectral density and coherence comparisons. Solar wind effects are not seen in a coherence plot between LPF and ACE; modest coherence in the planned LISA observational frequency band can be attributed to chance. This result indicates that measurable correlation due to solar-wind acceleration noise over 3 month timescales will be a negligible noise source. LISA is unlikely to inform solar wind measurements routinely. Another source of noise from the Sun, solar radiation pressure, is estimated to impart greater acceleration noise, but has yet to be analyzed.

79 ASTRONOMY AND ASTROPHYSICS↗

A Perspective on Data and Privacy for AI in Healthcare [Industrial and Governmental Activities]

As large language models continue to push the bounds of AI model size, they are also being trained on unprecedented volumes of data. While individual hospitals are estimated to produce petabytes of data per year, only a small fraction is currently being used for developing AI models. Additionally, with such data resources available, healthcare is well-positioned to benefit from the current trends in AI. Moreover, the inherently multi-modal and longitudinal nature of clinical data – from omics to imaging to unstructured notes – provides a fertile ground for the development and application of cutting-edge architectures like foundation models.

Gounley, John [Oak Ridge National Laboratory (ORNL↗

FY24 Progress Report: SRNL Analysis of ICCWR LCM and WAMS data for Corrosion and Cracking

Algorithms for Machine Learning (ML) and data analysis for the 3013 Surveillance Program have been developed in an ongoing collaborative effort by the Savannah River National Laboratory (SRNL) and the University of South Carolina (USC). The objective of the algorithms is to automate the identification of corrosion and crack formation in the Inner Container Closure Weld Region (ICCWR) of the canister system used to store Pu-bearing material. Data for corrosion and cracking is collected from large binary files generated by a Laser Confocal Microscope (LCM), the Wide Area 3D Measurement System (WAMS), or,in a recent proposal, by a Scanning Electron Microscope (SEM). The ML software uses the physical attributes in the data files (e.g., one or more of: height, color, and 16-bit grayscale values as functions of position in a plane projection) to detect signs of surface corrosion and cracking after being trained on similar data, with the features to be detected. Although the initial scope included screening for broader indicators of corrosion, e.g., pitting, identification of potential cracks was prioritized for the past several years at the request of program leadership. Labeled training data is essential to developing the ML algorithm, and enhancements to data labeling capability have been developed to address this essential precursor to application of ML routines. Efficient labeling is particularly important in view of the large volume of data required to train ML algorithms and the relative rarity of cracks in the ICCWR data set. The updated program will read binary data from either LCM, WAMS or SEM files, interrogate data attributes, facilitate user labeling of data for training ML algorithms, execute ML algorithms, output parameters from trained ML algorithms, report ML model accuracy with respect to labeled data, and generate graphical representations for various analyses. In FY24, hourglass neural networks (HNNs) that were initiated in FY22 were further developed and tested using available LCM data, and their performance was tested against that of the alternative U-Net Neural Network algorithm structure. HNNs along with previously developed Convolutional Neural Networks (CNNs) and Deep Neural Networks (DNNs) comprise a suite of ML tools for identification of cracks in the ICCWR

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

Poplar: a phylogenomics pipeline

Motivation Generating phylogenomic trees from the genomic data is essential in understanding biological systems. Each step of this complex process has received extensive attention and has been significantly streamlined over the years. Given the public availability of data, obtaining genomes for a wide selection of species is straightforward. However, analyzing that data to generate a phylogenomic tree is a multistep process with legitimate scientific and technical challenges, often requiring a significant input from a domain-area scientist. Results We present Poplar, a new, streamlined computational pipeline, to address the computational logistical issues that arise when constructing the phylogenomic trees. It provides a framework that runs state-of-the-art software for essential steps in the phylogenomic pipeline, beginning from a genome with or without an annotation, and resulting in a species tree. Running Poplar requires no external databases. In the execution, it enables parallelism for execution for clusters and cloud computing. The trees generated by Poplar match closely with state-of-the-art published trees. The usage and performance of Poplar is far simpler and quicker than manually running a phylogenomic pipeline. Availability and implementation Freely available on GitHub at https://github.com/sandialabs/poplar. Implemented using Python and supported on Linux.

Koning, Elizabeth [Sandia National Laboratories (S↗

AWSD Reactive Burn Model for the HMX‐Based High Explosive LX‐04

An Arrhenius–Wescott–Stewart–Davis (AWSD) reactive burn model is applied to describe shock initiation and detonation properties of the HMX-based high explosive LX-04. The parameters in the model are calibrated to data from multiple sources. The thermodynamic equations of state used in the model are calibrated to a combination of thermochemical calculations for HMX and LX-04 as well as experimentally-measured cylinder expansion results for LX-04. The kinetic parameters are calibrated to velocity data from gas gun experiments performed using EDC-32—a high explosive with the same chemical composition as LX-04 but different structural properties, and scaled rate stick data for PBX 9012. The AWSD model is shown to accurately describe the shock initiation and propagation of LX-04. Very good agreement is observed between the available experimental data and the AWSD model output. The presented results constitute an accurate LX-04 reactive burn model for use in engineering-scale models and simulations.

45 MILITARY TECHNOLOGY, WEAPONRY, AND NATIONAL DEF↗

Efficient data-driven regression for reduced-order modeling of spatial pattern formation

We present an efficient data-driven regression approach for constructing reduced-order models (ROMs) of reaction-diffusion systems exhibiting pattern formation. The ROMs are learned non-intrusively from available training data of physically accurate numerical simulations. The method can be applied to general nonlinear systems through the use of polynomial model form, while not requiring knowledge of the underlying physical model, governing equations, or numerical solvers. The process of learning ROMs is posed as a low-cost least-squares problem in a reduced-order subspace identified via Proper Orthogonal Decomposition (POD). Numerical experiments on classical pattern-forming systems–including the Schnakenberg and Mimura–Tsujikawa models–demonstrate that higher-order surrogate models significantly improve prediction accuracy while maintaining low computational cost. The proposed method provides a flexible, non-intrusive model reduction framework, well suited for the analysis of complex spatio-temporal pattern formation phenomena.

Data-driven modeling↗

Prospecting for Critical Minerals and Rare Earth Elements from Marcellus Shale in the Western Portion of the Appalachian Basin with Non-Destructive Core Characterization

Identification of sources for domestic critical minerals and rare earth elements (CM/REE) has been deemed essential for the energy transition by the United States Department of Energy (DOE). The U.S. DOE’s National Energy Technology Laboratory’s (NETL) Geomaterials Characterization Laboratory has performed non-destructive core characterizations on energy-relevant rock cores for the past decade. During this time, NETL has published over 36 technical reports and made the associated data publicly available. Much of this work focuses on unconventional shale gas, subsurface carbon storage systems, and carbon-ore. These efforts provide cm-scale petrophysical and elemental data, photographic documentation, detailed core descriptions, and computed tomography (CT) data for each well. This provides a first phase prospecting resource for CM/REE resources and can provide a map for pin-pointing intervals and lithologies for further development. Using historical core characterization data from 12 Marcellus wells from the western portion of the Appalachian Basin, this study builds an improved understanding of the chemostratigraphy of the basin. X-ray fluorescence (XRF) and CT images were used to determine lithologic intervals and potential ore bodies for further analysis, including benchtop digestion and inductively coupled plasma mass spectrometry (ICP-MS) to better understand the CM/REE enrichments.

Paronish, Thomas J.↗