Search NASA⌕ Search

SEARCH · Search NASA

Results for “Research data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 163 records · Page 9

FAIR Data Meets FAIR Software

Modern scientific research is increasingly defined by the interplay between data, software, and the workflows that connect them. Yet while the FAIR (Findable, Accessible, Interoperable, Reusable) principles have become foundational for scientific data stewardship, the same level of structure and expectation has only recently begun to extend to research software. This talk covers why and how FAIR principles are being applied to data and software to support data reuse. It outlines the gaps in current sharing norms, the growing federal emphasis on persistent identifiers and public access, and the opportunities created when datasets, computational workflows, code, and models are linked through rich, standardized metadata. Practical implementation pathways for the EIC and JLab communities are described, including datacards for structured dataset documentation and provenance-aware workflows. By aligning data lifecycle management with FAIR-aligned software practices, the scientific community can advance toward autonomous knowledge graphs, generative workflows, and high-quality, AI-ready scientific datasets.

McSpadden, Diana [Thomas Jefferson National Accele↗

DOE Data Days 2025 Report

The DOE Data Days (D3) workshop brings together data managers, developers, researchers, and program managers across the Department of Energy (DOE) and its national laboratories to highlight data management successes, identify potential synergies and common problems, and establish channels for collaboration across the DOE data management community.

97 MATHEMATICS AND COMPUTING↗

Physical, socio-psychological, and behavioural determinants of household energy consumption in the UK

Determining which attitudes and behaviours predict household energy consumption can help accelerate the low-carbon energy transition. Conventional approaches in this domain are limited, often relying on survey methods that produce data on individuals’ motivations and self-reported activities without pairing these with actual energy consumption records, which are particularly hard to collect for large, nationally representative samples. This challenge precludes the development of empirical evidence on which attitudes and behaviours influence patterns of energy consumption, thus limiting the extent to which these can inform energy interventions or conservation programs. This study demonstrates a novel methodology for estimating energy consumption in the absence of actual energy records by using a large, publicly available data set of energy consumption in the UK. We develop a predictive model using the Smart Energy Research Laboratory (SERL) data portal (with records from nearly 13,000 UK households) and then use this model to predict energy consumption (both electric and gas) for a sample of 1,000 UK householders for which we separately collect over 200 variables relating to climate change attitudes and practices. Our approach uses a set of over 50 independent variables that are shared between the data sets, allowing us to train a model on the SERL data and use it to analyse the relationship between energy consumption and the opinions, motivations, and daily practices of survey respondents. Results show that electricity consumption is influenced by a broader range of factors compared to gas. Household energy use is best explained by physical dwelling characteristics, socio-demographic variables, and certain behavioural and attitudinal measures. Notably, pro-environmental attitudes, frugality, and conscientiousness correlate with lower energy use, while income and consumerism are linked to higher consumption. We discuss how these findings can inform efforts to decarbonise home energy use in the UK.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Evaluating a Commercial Dynamic Line Rating Software with the National PMU Dataset

To accelerate the development of data-driven applications for power systems, the Department of Energy (DOE) supported the collection and curation of a synchrophasor dataset spanning two years of observations from transmission utilities across the US. This National PMU Dataset (NPDS) was anonymized and distributed to awardees of a DOE research grant under nondisclosure agreements (NDAs) but has also been retained at PNNL to enable further research. Agreements with data contributors prevent the data from being shared outside the organization. However, establishing a blind research validation methodology is envisioned to maximize the value proposition of the NPDS. In this validation strategy, researchers may share algorithms/software (potentially as executables to protect intellectual property) with PNNL, and PNNL will share feedback about the software’s performance on subsets of the NPDS. Such a blind methodology ensures that sensitive information about critical infrastructure remains protected, but the value of the NPDS can be extended to research beyond PNNL. Through iterative feedback, the algorithms may be tweaked to address real-world artifacts. As the NPDS data is temporally and geographically diverse, it may capture features absent in smaller datasets used during the development of the algorithm under test. This report presents lessons learned from applying the blind validation methodology to LineID™, a synchrophasor-based dynamic line rating software developed by Topolonet Corporation. Improvements made to the software through iterative feedback, limitations of the validation methodology, as well as how the limitations of the NPDS affected the evaluation process are discussed. Observations indicate that the proposed validation methodology can be valuable for evaluating other tools in the future.

97 MATHEMATICS AND COMPUTING↗

Deciphering spin-parity assignments of nuclear levels

Spin-parity assignments of nuclear levels are critical for understanding nuclear structure and reactions. However, inconsistent notation conventions and ambiguous reporting in research papers often lead to confusion and misinterpretations. Here, this paper examines the policies of the Evaluated Nuclear Structure Data File (ENSDF) and the evaluations by Endt and collaborators, highlighting key differences in their approaches to spin-parity notation. Sources of confusion are identified, including ambiguous use of strong and weak arguments and the conflation of new experimental results with prior constraints. Recommendations are provided to improve clarity and consistency in reporting spin-parity assignments, emphasizing the need for explicit notation conventions, clear differentiation of argument strengths, community education, and separate reporting of new findings. These steps aim to enhance the accuracy and utility of nuclear data for both researchers and evaluators.

Experimental Nuclear Physics↗

photoD with Rubin ’s Data Preview 1: First stellar photometric distances and faint blue star deficits

Aims. We investigate the utility of Rubin’s Data Preview 1 (DP1) for estimating stellar number density profiles across the Milky Way halo. Methods. We used stellar broad-band near-UV to near-IR ugrizy photometry released in Rubin’s DP1 to estimate distance and metallicity for blue main sequence stars brighter than r = 24 in three ~1.1 sq. deg. fields at southern Galactic latitudes. Results. Compared to TRILEGAL simulations of the Galaxy’s stellar content, we found a likely deficit of blue main sequence turn-off stars with 22 < r < 24. We interpreted this discrepancy as a signature of a steeper halo number density profile at galactocentric distances 10–50 kpc than the canonical ~1/r 3 profile assumed in TRILEGAL simulations. Conclusions. This interpretation is consistent with earlier suggestions based on observations of more luminous, but much less numerous, evolved stellar populations, along with a few pencil beam surveys of blue main sequence stars in the northern sky. These results bode well for the future Galactic halo exploration with Rubin’s Legacy Survey of Space and Time (LSST).

Galaxy: fundamental parameters↗

RTN-095: The Vera C. Rubin Observatory Data Preview 1

We present Rubin Data Preview 1 (DP1), the first release of data from the NSF-DOE Vera C. Rubin Observatory, consisting of raw and calibrated single-epoch images, coadds, difference images, detection catalogs, and other derived data products. DP1 is based on 1792 science-grade optical/near-infrared exposures acquired over 48 distinct nights by the Rubin Commissioning Camera, LSSTComCam, on the Simonyi Survey Telescope at the Summit Facility on Cerro Pachón, Chile during the first on-sky commissioning campaign in late 2024. DP1 covers a total of ~15 sq. deg. over seven roughly equally-sized non-contiguous fields, each independently observed in six broad photometric bands, ugrizy, spanning a range of stellar densities and latitudes and overlapping with external reference datasets. The median image quality across all bands, measured by the FWHM of the point-spread function, is approximately 1.13 arcseconds, with the sharpest images reaching about 0.65 arcseconds. DP1 contains approximately 2.3 million distinct astrophysical objects, of which 1.6 million are extended in at least one band, and 431 solar system objects, of which 93 are new discoveries. DP1 is approximately 3.5 TB in size and available to Rubin data rights holders via the Rubin Science Platform, a cloud-based environment for the analysis of petascale astronomical data. While small compared to future LSST releases, its high quality and diversity of data support a broad range of early science investigations across all four LSST themes, providing a valuable opportunity to engage with Rubin data ahead of the start of full operations in late 2025.

79 ASTRONOMY AND ASTROPHYSICS↗

Meteorological Services Annual Data Report for 2024

This document presents the meteorological data collected at Brookhaven National Laboratory (BNL) by Meteorological Services (Met Services) for the calendar year 2024. The purpose is to publicize the data sets available to emergency personnel, researchers and facility operations. Met services has been collecting data at BNL since 1949. Data from 1994 to the present is available in digital format. Data is presented in monthly plots of one-minute data. This allows the reader the ability to peruse the data for trends or anomalies that may be of interest to them. Full data sets are available to BNL personnel and to a limited degree outside researchers. The full data sets allow plotting the data on expanded time scales to obtain greater details (e.g., daily solar variability, inversions, etc.).

54 ENVIRONMENTAL SCIENCES↗

PV backsheets survey protocol: A framework for geo-spatial field surveys for bulk material characterization and reliability analysis applied across 41 PV systems

As widespread adoption of photovoltaic (PV) technologies continues, understanding the lifetime of modules is paramount to the viability of the industry as an environmentally conscious alternative to traditional energy generation. Although power degradation can affect the total energy production of a module over its lifetime, module safety failures necessitate the removal of a module leading to a loss of not only the particular asset, but the earning potential of the device. Therefore, it is critical to ensure that the components that provide essential safety functions for PV module operate for their entire rated lifetime. PV backsheets provide necessary electrical insulation to the completed device and failure of this component is cause for a immediate removal of the module. Degradation of the PV module backsheet has led to module safety failures in large-scale installations, costing millions of dollars in damages and lost potential revenue. The spatio-temporal degradation of fielded PV modules is important to study in order to identify which modules within installations are experiencing the greatest exposure conditions and in turn have the highest chance of failure. This paper describes a comprehensive field survey protocol developed for monitoring PV module backsheet performance using solely non-destructive methods in commercial PV fields. The protocol establishes a field naming convention, sampling method, data handling requirements, and measurement procedures. By ensuring consistent data collection practices, the field survey protocol enables research groups to obtain data of uniform quality on backsheet performance over multiple years and locations. In this study, the developed protocol was implemented at forty-one PV sites. Eight different types of airside layer backsheet materials including poly(vinylidene fluoride) (PVDF), acrylic PVDF, poly(tetrafluoroethylene-co-hexafluoropropylene-co-vinylidene fluoride) (THV), poly(vinyl fluoride) (PVF), poly(ethylene terephthalate) (PET), fluoroethylene vinyl ether (FEVE), polyethylene naphthalate (PEN), and glass were identified using attenuated total reflection Fourier transform infrared (ATR-FTIR) spectroscopy. The field survey results show that the spatial distribution of degradation indicators are non-uniform within a particular module, individual site, and across site locations. The degradation of PV modules increased in severity for modules mounted at the edge of rows (across a field) and near the junction box (within a module). This study demonstrates the sensitivity of material performance to exposure length across different materials and climates.

14 SOLAR ENERGY↗

Evaluation of daily gridded climate products using in situ FLUXNET data and tree growth modeling

Gridded climate data products have facilitated research in climate and ecology by providing meteorological data continuously across large spatial scales. However, the sensitivity of scientific outcomes to dataset choice remains poorly understood, and evaluation using station-based records can favor datasets built heavily on weather stations. Here, we evaluate seven high-resolution daily gridded datasets covering the contiguous United States using independent meteorology from the FLUXNET2015 dataset, with a focus on the implications of dataset choice for process-based tree growth modeling. We find that gridded products tend to capture temperature accurately while consistently overestimating the magnitude and frequency of precipitation and its extremes. Moreover, datasets vary in how they define a ‘day,’ which significantly affects temporal alignment with FLUXNET2015 observations. Despite differences among the datasets, the interannual variability in tree ring simulations is insensitive to dataset choice, likely because daily-scale biases are averaged out through accumulated growth across several months. However, inaccuracies in temperature and precipitation can significantly bias modeled xylem cell production, with systematically higher annual precipitation in the gridded datasets leading to greater xylem production compared to simulations using in situ data. Our results suggest that model applications, especially those that integrate to time scales longer than one day, are likely insensitive to climate dataset choice, but applications that are sensitive to daily climate variations or to absolute climate values need to carefully consider biases in gridded climate products.

54 ENVIRONMENTAL SCIENCES↗

Comprehensive framework for assessing and optimizing existing research networks

Conservation, monitoring, and research networks, or collections of ecological research sites unified under a common mission of data collection or a research mission, are essential infrastructure for understanding large landscapes. However, most networks developed opportunistically over decades rather than through systematic design, creating potential limitations in the ability to address conservation challenges across entire regions. We developed a framework to evaluate how well an existing research network represents the environmental conditions its members study and devised an approach to rank sites of priority for strategic expansion. Our approach measures performance through environmental representativeness, geographic coverage, and adequacy for scientific inference and thus optimizes limited monitoring resources to maximize scientific impact. We demonstrated this approach with the U.S. Department of Agriculture (USDA) Forest Service Experimental Forests and Ranges Network (EFRN), a 79‐site network across the United States that grew opportunistically over a century. At the national scale, the network effectively captured high‐biomass forests important for carbon cycle research; 82% of forest biomass was in well‐represented areas. Some areas in Texas, Florida, the Rocky Mountains, and the West Coast had no relevant EFRN sites, which limits the ability to make regional inferences. A fundamental challenge for the EFRN was that sites improving regional extent coverage sometimes provided minimal national benefits, which can create conflicts between local and global priorities. Adding the highest‐ranked candidate site provided a relevant site for 17% of currently poorly represented 1‐km pixel cells nationally, but regional and national site rankings varied considerably due to nested spatial inference. This framework provides quantitative tools for strategic infrastructure decision‐making, ensures that limited monitoring resources maximize conservation impact, and can be applied broadly to address the widespread challenge of optimizing conservation and monitoring networks worldwide.

additional site↗

Analysis on Evaluations of Monterey Bay Aquarium Research Institute’s Wave Energy Converter’s Field Data Using WEC-Sim and Gazebo: A Simulation Tool Comparison

Although many studies have validated wave energy converter (WEC) numerical models against scaled prototype experimental data, there remains a notable lack of validation using data from full-scale deployed WECs. This paper compares two numerical models of Monterey Bay Aquarium Research Institute’s Wave Energy Converter (MBARI-WEC), a two-body point absorber with an electro-hydraulic power take-off system (PTO). The models are implemented in WEC-Sim/Simscape and Gazebo Simulator. A statistical analysis of the models was performed, and field results were obtained to compare the models’ accuracy in predicting the RMS piston velocity, RMS motor speed, and mean electric power compared to field data for 56 observations across varying sea states. The Gazebo model demonstrated a closer agreement across all three parameters for a majority of the observations. When compared to the field data, the Gazebo and WEC-Sim models exhibited average mean electric power overestimations of 13% and 22%, respectively.

16 TIDAL AND WAVE POWER↗

The Vera C. Rubin Observatory Data Preview 1

We present Rubin Data Preview 1 (DP1), the first data from the National Science Foundation–Department of Energy Vera C. Rubin Observatory, comprising raw and calibrated single-epoch images, coadds, difference images, detection catalogs, and ancillary data products. DP1 is based on 1792 optical–near-infrared exposures acquired over 48 distinct nights by the Rubin Commissioning Camera (LSSTComCam) on the Simonyi Survey Telescope at the Summit Facility on Cerro Pachón, Chile in late 2024. DP1 covers ∼15 deg 2 distributed across seven roughly equal-sized noncontiguous fields, each independently observed in six broad photometric bands, ugrizy. The median FWHM of the point-spread function across all bands is approximately 1"14, with the sharpest images reaching about 0." 58. The 5σ point-source depths for coadded images in the deepest field, the Extended Chandra Deep Field South, are u = 24.55, g = 26.18, r = 25.96, i = 25.71, z = 25.07, and y = 23.1. Other fields are no more than 2.2 mag shallower in any band, where they have nonzero coverage. DP1 contains approximately 2.3 million distinct astrophysical objects, of which 1.6 million are extended in at least one band in coadds, and 431 solar system objects, of which 93 are new discoveries. DP1 is approximately 3.5 TB in size and is available to Vera C. Rubin Observatory data rights holders via the Rubin Science Platform, a cloud-based environment for the analysis of petascale astronomical data. While small compared to future LSST releases, its high quality and diversity of data support a broad range of early science investigations ahead of full operations in 2026.

Ground-based astronomy↗

Supporting Data for "Constraints on magnetism and correlations in RuO2 from lattice dynamics and Mössbauer spectroscopy"

Contents of this DOI are data for the research paper "Constraints on magnetism and correlations in RuO2 from lattice dynamics and Mossbauer spectroscopy" by the authors: George Yumnam, et. al. The dataset contains data from inelastic x-ray scattering measurements of RuO2 single crystals performed at the Advanced Photon Source in Argonne National Lab. This dataset also contains data from inelastic neutron scattering experiments of RuO2 powder performed at the ARCS spectrometer at Spallation Neutron Source at ORNL. Supporting theoretical calculations based on density functional theory with r2SCAN and DFT+U methods is also included in this dataset. Please see the README.txt files for more information of the dataset and file contents. For comments and questions, please contact: George Yumnam (yumnamg@ornl.gov) and/or Raphael P Hermann (hermannrp@ornl.gov)

density functional theory↗

PSTN-019: The LSST Science Pipelines Software: Optical Survey Pipeline Reduction and Analysis Environment

The NSF-DOE Vera C. Rubin Observatory is executing the Legacy Survey of Space and Time (LSST) as its prime mission, producing a series of data releases over the ten-year survey. The LSST Science Pipelines Software will be used to create these data releases and to perform the nightly prompt processing and alert production. This paper provides an overview of the LSST Science Pipelines Software, describing the components and their integration into pipelines that generate science-ready data products.

79 ASTRONOMY AND ASTROPHYSICS↗

Supporting Data for "On the magnetic contribution of itinerant electrons to neutron diffraction in the topological antiferromagnet CeAlGe"

Contents of this DOI are data for the research paper "On the magnetic contribution of itinerant electrons to neutron diffraction in the topological antiferromagnet CeAlGe" by the authors: V. Pomjakushin et al. The dataset contains raw data from neutron scattering experiments of CeAlGe powder performed at the CNCS spectrometer at Spallation Neutron Source at ORNL.

36 MATERIALS SCIENCE↗

A Bayesian Learning Approach to Wireless Outdoor Heatmap Construction using Deep Gaussian Process

We present a novel Bayesian learning approach to outdoor radio heatmap construction utilizing deep Gaussian process (GP). The proposed approach employs a two-layer hierarchy which consists of two cascaded Gaussian processes that are capable of modeling more complex input-output relations than standard single-layer Gaussian processes. Since deriving the exact model likelihood is challenging, a lower bound is optimized instead so that gradient descent-based methods can be performed to find out the optimal model parameters. Typically, inducing points are used in GPs to facilitate low-rank approximation of covariance (kernel) matrices for computation speedup. However, the inaccuracy induced by inducing points can accumulate when stacking multiple layers of GP which may hinder the performance of deep GP. Moreover, since inducing points need to be learned, having them at all layers of deep GP also incurs computational burden. To overcome the above challenges, in contrast to the canonical deep GP model, we use a modified architecture where a full standard GP resides in the first layer and inducing points are only introduced for the second layer. This modified architecture strikes a balance between model accuracy and training complexity. In the proposed model, the noise parameter of the first GP layer is also eliminated to improve the training efficiency as the noise parameter at the output of the second layer suffices to model the uncertainty in the output. The proposed approach is evaluated on real-world datasets, in the form of location-Received Signal Strength (RSS) pairs, collected from the Platform for Open Wireless Data-driven Experimental Research (POWDER) located at the campus of the University of Utah. Experiment results show that the proposed approach can achieve smaller prediction errors on various training and testing data configurations than DNN-based and GP-based methods.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

rcsb-api : Python Toolkit for Streamlining Access to RCSB Protein Data Bank APIs

The Protein Data Bank (PDB) was founded in 1971 as the first open-access digital data resource in biology to serve as the single global archive for three-dimensional (3D) macromolecular structure data. Current PDB holdings exceed 230,000 experimentally determined structures of proteins, nucleic acids, viruses, and macromolecular machines. The RCSB Protein Data Bank RCSB.org research-focused web portal facilitates search, analyses, and visualization of every PDB structure along with more than one million Computed Structure Models from AlphaFold DB and the ModelArchive. It is powered by a set of publicly available Application Programming Interfaces (APIs) that both support RCSB.org users and provide programmatic access to PDB data. Given the breadth and levels of granularity encompassed in this rich data collection, efficiently accessing the information programmatically may be challenging for new users. RCSB PDB has developed a Python software package, rcsb-api , that facilitates easy and efficient use of RCSB PDB APIs within a Python environment. This software tool is designed to streamline access to the extensive corpus of data housed within the PDB, enabling researchers to search, retrieve, and analyze 3D biostructure data seamlessly. Its use will accelerate research in structural biology, molecular biology and biochemistry, drug discovery, and bioinformatics by providing more efficient tools for data integration and analysis. The new toolkit is available on GitHub (github.com/rcsb/py-rcsb-api) and published to the public Python package repository (PyPI) to foster wider usage and support basic and applied research in fundamental biology, biomedicine, and the energy sciences.

FAIR principles↗