Search NASA⌕ Search

SEARCH · Search NASA

Results for “Data exploration”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

RE Data Explorer is Expanding Public Access to High-Quality Data and Analytics

The Renewable Energy (RE) Data Explorer is an innovative and intuitive web platform from the U.S. Agency for International Development (USAID)-National Renewable Energy Laboratory (NREL) Partnership that enables users to easily access high-quality renewable energy resource data and applied analytics.

ENERGY PLANNING, POLICY, AND ECONOMY↗

Basin and Range Investigation for Developing Geothermal Energy: Exploration Data

This data package includes exploration material from the Basin & Range Investigation for Developing Geothermal Energy [in Hidden Systems] project (BRIDGE), which is part of a broader initiative to advance the exploration of hidden geothermal resources in the Basin & Range Province of the western U.S. Data modalities include a helicopter-borne time-domain electromagnetic survey, magnetotellurics, 2-meter temperature measurements, ground-based gravity and legacy aeromagnetic surveys, geochemistry, geologic mapping, LiDAR analysis, 3D models, associated geospatial data, and a bibliography of existing data and references utilized in prospect characterization and conceptual modeling. Key files are in CSV, Geosoft, and Geotools formats. Please refer to READMEs for dataset-specific information. Where applicable, acquisition data and inversion models for a particular prospect or area of interest are organized separately. This BRIDGE data package is the product of a collaboration led by Sandia National Laboratories with partners from Geologica Geothermal Group, Inc., the U.S. Navy Geothermal Program Office, and consultants Steven Sewell (Australis Geoscience Ltd) and William Cumming (Cumming Geoscience). The project's areas of interest (AOIs) are based off priority areas of interest in the southwestern portion of the Nevada Play Fairway map, distribution across tectonic provinces, accessibility, and the project team's extensive experience in the region. AOIs cover about a dozen basins that include unexplored prospects, partially explored prospects, and some developed analogue resources that provide validation cases. Many unexplored and partially explored prospects are on U.S. Department of Defense (DoD) land, though adjacent lands are included as well.

15 GEOTHERMAL ENERGY↗

Mining G.O.L.D. (Geothermal Opportunities Leveraged Through Data): Exploring Synergies Between the Geothermal and Mining Industries

This report analyzes potential collaborations between the geothermal and locatable mineral industries (focused on the portion of the Basin and Range Province within Nevada, United States of America). The objectives of this study included analyzing: 1. The type and quality of data collected by the locatable mineral industry to determine feasibility for geothermal resource exploration; 2. The regulatory pathways and potential barriers that could prevent development of geothermal resources discovered via a mining claim (and vice versa) in the United States; 3. The historical development of geothermal resources discovered via mineral exploration data in the United States and illustrations of co-located mining and geothermal power projects; 4. The value propositions for both the locatable mineral and the geothermal industries to collaborate. This article concludes that many of the data collected by the mining industry as part of locatable mineral exploration (e.g., copper, gold, lithium) would also be useful for identifying and developing previously unknown geothermal resources (and in some notable cases, already have led to geothermal resource development). In addition, for minimal costs, the mining industry could catalogue these data and potentially monetize the data itself or use the data in the future to develop a geothermal project. Leveraging these locatable mineral data to develop geothermal resources and/or co-located minerals and geothermal resources would represent significant cost savings when compared to developing geothermal resources under a business-as-usual scenario as well as compared to current generating technologies (e.g., diesel-powered generators) employed at remote mining operations. Ultimately, leveraging mining industry data, knowledge, and expertise serves to effectively expand the geothermal exploration workforce, increase the rate of geothermal resource discovery, and potentially reduce geothermal electricity's levelized cost of energy (LCOE) by 23%-29%.

15 GEOTHERMAL ENERGY↗

NeuralCubes: Deep Representations for Visual Data Exploration

Visual exploration of large multi-dimensional datasets has seen tremendous progress in recent years, allowing users to express rich data queries that produce informative visual summaries, all in real time. Techniques based on data cubes are some of the most promising approaches. However, these techniques usually require a large memory footprint for large datasets. To tackle this problem, we present NeuralCubes: neural networks that predict results for aggregate queries, similar to data cubes. NeuralCubes learns a function that takes as input a given query, for instance, a geographic region and temporal interval, and outputs the result of the query. The learned function serves as a real-time, low-memory approximator for aggregation queries. Our models are small enough to be sent to the client side (e.g. the web browser for a web-based application) for evaluation, enabling data exploration of large datasets without database/network connection. Here, we demonstrate the effectiveness of NeuralCubes through extensive experiments on a variety of datasets and discuss how NeuralCubes opens up opportunities for new types of visualization and interaction.

97 MATHEMATICS AND COMPUTING↗

Utilizing Novel Field and Data Exploration Methods to Explore Hot Moments in High-Frequency Soil Nitrous Oxide Emissions Data: Opportunities and Challenges

Soil nitrous oxide (N 2 O) emissions are an important driver of climate change and are a major mechanism of labile nitrogen (N) loss from terrestrial ecosystems. Evidence increasingly suggests that locations on the landscape that experience biogeochemical fluxes disproportionate to the surrounding matrix (hot spots) and time periods that show disproportionately high fluxes relative to the background (hot moments) strongly influence landscape-scale soil N 2 O emissions. However, substantial uncertainties remain regarding how to measure and model where and when these extreme soil N 2 O fluxes occur. High-frequency datasets of soil N 2 O fluxes are newly possible due to advancements in field-ready instrumentation that uses cavity ring-down spectroscopy (CRDS). Here, we outline the opportunities and challenges that are provided by the deployment of this field-based instrumentation and the collection of high-frequency soil N 2 O flux datasets. While there are substantial challenges associated with automated CRDS systems, there are also opportunities to utilize these near-continuous data to constrain our understanding of dynamics of the terrestrial N cycle across space and time. Finally, we propose future research directions exploring the influence of hot moments of N 2 O emissions on the N cycle, particularly considering the gaps surrounding how global change forces are likely to alter N dynamics in the future.

54 ENVIRONMENTAL SCIENCES↗

Faraday: A High-temperature Electrolysis Data Explorer

Faraday is a high-temperature electrolysis data visualization tool, which reveals the performance of various button cells under test conditions. These tests and the resulting analytics on their data constitute a state of the industry as the US Department of Energy pushes for the production of hydrogen. Faraday leverages the Idaho National Laboratory's DeepLynx data warehouse to standardize and query button cell data. Faraday programmatically accesses this data in DeepLynx by traversing the schema, represented by a custom ontology. The user interface queries DeepLynx for timeseries data associated with specific button cells in the warehouse, and renders them using JavaScript charts. Additional charting and data analysis techniques are made possible by an auxiliary Python server.

Woodruff, Nathan↗

Analysis of Selected Publicly Available Geothermal Exploration Data Gaps

As part of a United States Department of Energy (DOE) supported retrospective analysis of DOE's Play Fairway Analysis (PFA) projects, the National Renewable Energy Laboratory (NREL) compiled and analyzed publicly available geothermal exploration datasets to identify and highlight data gaps in areas prospective for hosting geothermal resources. The analysis was intended to understand the existing geographic coverage of selected datasets commonly utilized both by the PFA projects and geothermal developers during resource assessments including geologic mapping, temperature gradient drilling, and aeromagnetic, gravimetric, and lidar surveys. Results indicate that broad areas of the western United States estimated to have geothermal potential lack sufficient geologic and geophysical coverage necessary for even regional resource exploration. The study directly informed the recent Geoscience Data Acquisition for Western Nevada, or GeoDAWN - which united DOE's Geothermal Technologies Office (GTO) with the U.S. Geological Survey (USGS) of the U.S. Department of the Interior to assist U.S. needs for energy and critical minerals. The study also has the potential to inform public investment in further data acquisition for characterization of the Earth both for geothermal and other natural resource assessments.

data↗

High-throughput predictions of metal–organic framework electronic properties: theoretical challenges, graph neural networks, and data exploration

Abstract With the goal of accelerating the design and discovery of metal–organic frameworks (MOFs) for electronic, optoelectronic, and energy storage applications, we present a dataset of predicted electronic structure properties for thousands of MOFs carried out using multiple density functional approximations. Compared to more accurate hybrid functionals, we find that the widely used PBE generalized gradient approximation (GGA) functional severely underpredicts MOF band gaps in a largely systematic manner for semi-conductors and insulators without magnetic character. However, an even larger and less predictable disparity in the band gap prediction is present for MOFs with open-shell 3 d transition metal cations. With regards to partial atomic charges, we find that different density functional approximations predict similar charges overall, although hybrid functionals tend to shift electron density away from the metal centers and onto the ligand environments compared to the GGA point of reference. Much more significant differences in partial atomic charges are observed when comparing different charge partitioning schemes. We conclude by using the dataset of computed MOF properties to train machine-learning models that can rapidly predict MOF band gaps for all four density functional approximations considered in this work, paving the way for future high-throughput screening studies. To encourage exploration and reuse of the theoretical calculations presented in this work, the curated data is made publicly available via an interactive and user-friendly web application on the Materials Project.

36 MATERIALS SCIENCE↗

Exploring Data Set Bias and Decision Support with Predictive Uncertainty Through Bayesian Approximations and Convolutional Neural Networks

Individual seismic catalogs can contain multiscale observations from fault level to global scales and associated waveforms from discrete events reflect crustal structure across many different scales and locations. Seismic network aperture, geographic location, and observation distance may not provide informative guidance or intuition on how different catalogs will behave across models trained under different conditions. We rely on uncertainty to provide guardrails for when to trust model decisions, but understanding when our uncertainty is trustworthy is an open challenge. Here, in this work, we explore Bayesian approximation methods for assigning predictive uncertainty in seismic event classification problems. We find that computationally expensive Bayesian approximations do not outperform simple ensemble methods. We also find that when exploiting multiple seismic event catalogs, joint training with data from all the catalogs combined with Bayesian approximations and supervised training for classification can obscure bias and result in less robust uncertainty while also not providing substantial performance benefits compared to training individual models for each catalog.

58 GEOSCIENCES↗

47 Tuc in Rubin Data Preview 1. Exploring Early LSST Data and Science Potential

We present analyses of the early data from Rubin Observatory’s Data Preview 1 (DP1) for the field of the globular cluster 47 Tuc. The DP1 data set for 47 Tuc includes four nights of observations from the Rubin Commissioning Camera (LSSTComCam), covering multiple bands (ugriy). We address challenges of crowding in the inner region of the cluster and toward the SMC in DP1, and demonstrate improved star–galaxy separation by fitting fifth-degree polynomials to the stellar loci in color–color diagrams and applying multidimensional sigma clipping. We compile a catalog of 3576 probable 47 Tuc member stars selected via a combination of isochrone, Gaia proper-motion, and color–color space matched filtering. We explore the sources of photometric scatter in the 47 Tuc color–color sequence, evaluating contributions from various potential sources, including differential extinction within the cluster. Finally, of the 72 well-characterized variables in the field, we recover three known variable stars, including two RR Lyrae and one eclipsing binary, in the coadd-based object catalog, and identify 62 in the difference image-based object catalog. Although the DP1 lightcurves have sparse temporal sampling, they appear to follow the patterns of densely sampled literature lightcurves well. Despite some data limitations for crowded-field stellar analysis, DP1 demonstrates the promising scientific potential for future LSST data releases.

Choi, Yumi [NSF National Optical-Infrared Astronom↗

Statistical Validation of Multiple Related Data Sets—Case Study Using Interstellar Boundary Explorer Satellite Data

Abstract Space scientists often face the question of whether data collected by different instruments are measurements of the same source population. This paper proposes a statistical validation method for evaluating the agreement between such related data sets. It offers a detailed case study focused on validating a new data set from the Interstellar Boundary Explorer (IBEX) mission, which serves as a practical how-to guide for similar analyses. Since 2008, the IBEX satellite has been gathering data on heliospheric energetic neutral atoms (ENAs) while being exposed to various sources of background noise, such as cosmic rays and solar energetic particles. The IBEX mission initially released only a qualified triple-coincidence (qABC) data product, which was designed to provide observations of ENAs free of background contamination. Further measurements revealed that the qABC data were in fact susceptible to contamination, having relatively low ENA counts and high background rates. To mitigate this issue, the mission team recently considered releasing a certain qualified double-coincidence (qBC) data product, which has roughly twice the detection rate of the qABC data product. This paper presents a simulation-based validation of the new qBC data product against the already-released qABC data product. The results show that the qBCs can plausibly be said to be measuring the same source population as the qABCs up to an average absolute deviation of 3.6%. Visual diagnostics provide additional confirmation of source rate coherence across data products. The framework introduced here is general and can be applied to other validation problems both within and outside the field of space physics.

79 ASTRONOMY AND ASTROPHYSICS↗

Identifying genomic data use with the Data Citation Explorer

Increases in sequencing capacity, combined with rapid accumulation of publications and associated data resources, have increased the complexity of maintaining associations between literature and genomic data. As the volume of literature and data have exceeded the capacity of manual curation, automated approaches to maintaining and confirming associations among these resources have become necessary. Here we present the Data Citation Explorer (DCE), which discovers literature incorporating genomic data that was not formally cited. This service provides advantages over manual curation methods including consistent resource coverage, metadata enrichment, documentation of new use cases, and identification of conflicting metadata. The service reduces labor costs associated with manual review, improves the quality of genome metadata maintained by the U.S. Department of Energy Joint Genome Institute (JGI), and increases the number of known publications that incorporate its data products. The DCE facilitates an understanding of JGI impact, improves credit attribution for data generators, and can encourage data sharing by allowing scientists to see how reuse amplifies the impact of their original studies.

59 BASIC BIOLOGICAL SCIENCES↗

Data Citation Explorer (DCE) v1.0

Increases in sequencing capacity, combined with rapid accumulation of publications and associated data resources, have increased the complexity of maintaining associations between literature and genomic data. As the volume of literature and data have exceeded the capacity of manual curation, automated approaches to maintaining and confirming associations among these resources have become necessary. Here we present the Data Citation Explorer (DCE), which discovers literature incorporating genomic data whether or not provenance was clearly indicated. This service provides advantages over manual curation methods including consistent resource coverage, metadata enrichment, documentation of new use cases, and identification of conflicting metadata. The service reduces labor costs associated with manual review, improves the quality of genome metadata maintained by the U.S. Department of Energy Joint Genome Institute (JGI), and increases the number of known publications that incorporate its data products. The DCE facilitates an understanding of JGI impact, improves credit attribution for data generators, and can encourage data sharing by allowing scientists to see how reuse amplifies the impact of their original studies.

Parker, Charles↗

OpenET: Filling a Critical Data Gap in Water Management for the Western United States

The lack of consistent, accurate information on evapotranspiration (ET) and consumptive use of water by irrigated agriculture is one of the most important data gaps for water managers in the western United States (U.S.) and other arid agricultural regions globally. The ability to easily access information on ET is central to improving water budgets across the West, advancing the use of data-driven irrigation management strategies, and expanding incentive-driven conservation programs. Recent advances in remote sensing of ET have led to the development of multiple approaches for field-scale ET mapping that have been used for local and regional water resource management applications by U.S. state and federal agencies. The OpenET project is a community-driven effort that is building upon these advances to develop an operational system for generating and distributing ET data at a field scale using an ensemble of six well-established satellite-based approaches for mapping ET. Key objectives of OpenET include: Increasing access to remotely sensed ET data through a web-based data explorer and data services; supporting the use of ET data for a range of water resource management applications; and development of use cases and training resources for agricultural producers and water resource managers. Here we describe the OpenET framework, including the models used in the ensemble, the satellite, meteorological, and ancillary data inputs to the system, and the OpenET data visualization and access tools. We also summarize an extensive intercomparison and accuracy assessment conducted using ground measurements of ET from 139 flux tower sites instrumented with open path eddy covariance systems. Results calculated for 24 cropland sites from Phase I of the intercomparison and accuracy assessment demonstrate strong agreement between the satellite-driven ET models and the flux tower ET data. For the six models that have been evaluated to date (ALEXI/DisALEXI, eeMETRIC, geeSEBAL, PT-JPL, SIMS, and SSEBop) and the ensemble mean, the weighted average mean absolute error (MAE) values across all sites range from 13.6 to 21.6 mm/month at a monthly timestep, and 0.74 to 1.07 mm/day at a daily timestep. At seasonal time scales, for all but one of the models the weighted mean total ET is within ±8% of both the ensemble mean and the weighted mean total ET calculated from the flux tower data. Overall, the ensemble mean performs as well as any individual model across nearly all accuracy statistics for croplands, though some individual models may perform better for specific sites and regions. We conclude with three brief use cases to illustrate current applications and benefits of increased access to ET data, and discuss key lessons learned from the development of OpenET.

54 ENVIRONMENTAL SCIENCES↗

simple_RDE_model

Jupyter notebook containing a simple rotating detonation engine model to generate synthetic data for exploring data analysis techniques.

42 ENGINEERING↗

Methods for rapid identification of anomalous layers in laser powder bed fusion

In situ process monitoring is a key requirement for increased industry acceptance of powder bed Additive Manufacturing. As sensing technologies increase in maturity, attention must also be given to effective data exploration techniques. These data are often high-resolution and multi-modal, with each build consisting of thousands of layers. Here, in this paper, the authors propose two methods enabling users to rapidly identify layers of interest within a build. Both methods leverage results from deep learning based segmentations of in situ powder bed images. The first method is an unsupervised “reverse layer search” algorithm while the second method uses supervised machine learning.

36 MATERIALS SCIENCE↗

The DECam Local Volume Exploration Survey Data Release 2

We present the second public data release (DR2) from the DECam Local Volume Exploration survey (DELVE). DELVE DR2 combines new DECam observations with archival DECam data from the Dark Energy Survey, the DECam Legacy Survey, and other DECam community programs. DELVE DR2 consists of ∼160,000 exposures that cover >21,000 deg$^{2}$ of the high-Galactic-latitude (∣b∣ > 10°) sky in four broadband optical/near-infrared filters (g, r, i, z). DELVE DR2 provides point-source and automatic aperture photometry for ∼2.5 billion astronomical sources with a median 5σ point-source depth of g = 24.3, r = 23.9, i = 23.5, and z = 22.8 mag. A region of ∼17,000 deg$^{2}$ has been imaged in all four filters, providing four-band photometric measurements for ∼618 million astronomical sources. DELVE DR2 covers more than 4 times the area of the previous DELVE data release and contains roughly 5 times as many astronomical objects. DELVE DR2 is publicly available via the NOIRLab Astro Data Lab science platform.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗