Search NASA⌕ Search

SEARCH · Search NASA

Results for “Dataset”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 523 records · Page 29

High-resolution fully-polarimetric synthetic aperture radar dataset

Fully-polarimetric synthetic aperture radar (PolSAR) data contain a rich body of elementary scattering physics information that is critically valuable for a broad range of applications and scientific purposes. However, there is a lack of available high-resolution (< 0.3048-m) data available for PolSAR phenomenology research. This article introduces a high-resolution PolSAR data set collected and provided by Sandia National Laboratories (SNL). The data sets were collected to support studying high-resolution scattering physics from different types of clutter and applications such as polarimetric-based terrain classification.

West, Roger Derek↗

Sea cucumber ( Holothuria glaberrima ) intestinal microbiome dataset from Puerto Rico, generated by shotgun sequencing

The sea cucumber (H. glaberrima) is a species found in the shallow waters near coral reefs and seagrass beds in Puerto Rico. To characterize the microbial taxonomic composition and functional profiles present in the sea cucumber, total DNA was obtained from their intestinal system, fosmid libraries constructed, and subsequent sequencing was performed. The diversity profile displayed that the most predominant domain was Bacteria (76.56 %), followed by Viruses (23.24 %) and Archaea (0.04 %). Within the 11 phyla identified, the most abundant was Proteobacteria (73.16 %), followed by Terrabacteria group (3.20 %) and Fibrobacterota, Chlorobiota, Bacteroidota (FCB) superphylum (1.02 %). The most abundant species were Porvidencia rettgeri (21.77 %), Pseudomonas stutzeri (14.78 %), and Alcaligenes faecalis (5.00 %). The functional profile revealed that the most abundant functions are related to transporters, MISC (miscellaneous information systems), organic nitrogen, energy, and carbon utilization. The data collected in this project on the diversity and functional profiles of the intestinal system of the H. glaberrima provided a detailed view of its microbial ecology. These findings may motivate comparative studies aimed at understanding the role of the microbiome in intestinal regeneration.

59 BASIC BIOLOGICAL SCIENCES↗

Dataset of mechanically induced thermal runaway measurement and severity level on Li-ion batteries

The deployment of Li-ion batteries covers a wide range of energy storage applications, from mobile phones, e-bikes, electric vehicles (EV) and stationary energy storage systems. However, safety issue such as thermal runaway is always one of the most important concerns to prevent Li-ion batteries from further market penetration. A standardized single-side indentation test protocol was developed to mechanically induce an internal short-circuit. The cell voltage, compressive load, indenter stroke, and temperature at the indentation point are measured in time series. The test data of each cell, along with cell parameters such as dimensions, mass, chemistry, state of charge (SOC), capacity, are integrated together to calculate a thermal runaway severity score from 0 to100. Complete data collection process including the original measured record, test method, severity score calculation scheme is presented in this article. The thermal runaway severity analysis and the more than 100 tested Li-ion battery records provide a good data source for further comparison and ranking of thermal runaway risks.

25 ENERGY STORAGE↗

The Open DAC 2023 Dataset and Challenges for Sorbent Discovery in Direct Air Capture

Direct air capture (DAC) of CO 2 with porous adsorbents such as metal-organic frameworks (MOFs) has the potential to aid large-scale decarbonization. Previous screening of MOFs for DAC relied on empirical force fields and ignored adsorbed H 2 O and MOF deformation. We performed quantum chemistry calculations overcoming these restrictions for thousands of MOFs. The resulting data enable efficient descriptions using machine learning.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Non-linear relationships between daily temperature extremes and US agricultural yields uncovered by global gridded meteorological datasets

Global agricultural commodity markets are highly integrated among major producers. Prices are driven by aggregate supply rather than what happens in individual countries in isolation. Furthermore, estimating the effects of weather-induced shocks on production, trade patterns and prices hence requires a globally representative weather data set. Recently, two data sets that provide daily or hourly records, GMFD and ERA5-Land, became available. Starting with the US, a data rich region, we formally test whether these global data sets are as good as more fine-scaled country-specific data in explaining yields and whether they estimate similar response functions. While GMFD and ERA5-Land have lower predictive skill for US corn and soybeans yields than the fine-scaled PRISM data, they still correctly uncover the underlying non-linear temperature relationship. All specifications using daily temperature extremes under any of the weather data sets outperform models that use a quadratic in average temperature. Correctly capturing the effect of daily extremes has a larger effect than the choice of weather data. In a second step, focusing on Sub Saharan Africa, a data sparse region, we confirm that GMFD and ERA5-Land have superior predictive power to CRU, a global weather data set previously employed for modeling climate effects in the region.

54 ENVIRONMENTAL SCIENCES↗

A mapped dataset of surface ocean acidification indicators in large marine ecosystems of the United States

Mapped monthly data products of surface ocean acidification indicators from 1998 to 2022 on a 0.25° by 0.25° spatial grid have been developed for eleven U.S. large marine ecosystems (LMEs). The data products were constructed using observations from the Surface Ocean CO 2 Atlas, co-located surface ocean properties, and two types of machine learning algorithms: Gaussian mixture models to organize LMEs into clusters of similar environmental variability and random forest regressions (RFRs) that were trained and applied within each cluster to spatiotemporally interpolate the observational data. The data products, called RFR-LMEs, have been averaged into regional timeseries to summarize the status of ocean acidification in U.S. coastal waters, showing a domain-wide carbon dioxide partial pressure increase of 1.4 ± 0.4 μatm yr -1 and pH decrease of 0.0014 ± 0.0004 yr -1 . RFR-LMEs have been evaluated via comparisons to discrete shipboard data, fixed timeseries, and other mapped surface ocean carbon chemistry data products. Regionally averaged timeseries of RFR-LME indicators are provided online through the NOAA National Marine Ecosystem Status web portal.

54 ENVIRONMENTAL SCIENCES↗

Energy dataset of Frontier supercomputer for waste heat recovery

The Hewlett Packard Enterprise–Cray EX Frontier is the world’s first and fastest exascale supercomputer, hosted at the Oak Ridge Leadership Computing Facility in Tennessee, United States. Frontier is a significant electricity consumer, drawing 8–30 MW; this massive energy demand produces significant waste heat, requiring extensive cooling measures. Although harnessing this waste heat for campus heating is a sustainability goal at Oak Ridge National Laboratory (ORNL), the 30 °C–38 °C waste heat temperature poses compatibility issues with standard HVAC systems. Heat pump systems, prevalent in residential settings and some industries, can efficiently upgrade low-quality heat to usable energy for buildings. Thus, heat pump technology powered by renewable electricity offers an efficient, cost-effective solution for substantial waste heat recovery. However, a major challenge is the absence of benchmark data on high-performance computing (HPC) heat generation and waste heat profiles. This paper reports power demand and waste heat measurements from an ORNL HPC data centre, aiming to guide future research on optimizing waste heat recovery in large-scale data centres, especially those of HPC calibre.

97 MATHEMATICS AND COMPUTING↗

Random forest models accurately classify synthetic opioids using high-dimensionality mass spectrometry datasets

Detection of novel threat agents presents several challenges, a principle one being the development of untargeted methods to screen an increasing number of threat chemicals whose exact structures are unknown. With the use of Machine Learning (ML) tools, we can guide the development of analytical methods for broad-spectrum detection of unbounded threat chemical families in complex mixtures. Toward this goal, we used nominal mass and high-resolution mass spectrometry data for hundreds of synthetic opioids and non-opioid compounds. We tested two ML techniques, logistic regression and random forest, to develop models towards a practical, implementable method for opioid detection. We found that of these tested ML methods, random forest models resulted in the highest validation accuracy (95+%) for both nominal mass and high-resolution classification of opioids versus non-opioids, with low false positive and false negative rates. The RF models were then used to successfully predict the classification of 10 compounds—five opioids and five non-opioids not part of the training and validation analysis. This application of ML is a critical step towards the development of field-deployable nominal mass spectrometers with ML-driven analyses for classification of emergent threats.

Chemistry↗