Search NASA⌕ Search

SEARCH · Search NASA

Results for “data access”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 253 records · Page 14

MINERvA Experiment Open Data Release

This is an open data product that contains all the neutrino and antineutrino data from the MINERvA experiment, packaged in a way that you can use. The ultimate package will contain both neutrino and antineutrino data, and will contain both our Low energy and Medium Energy data, where the neutrino energy distributions peaked around 3GeV and 6GeV respectively. We also provide simulated data and a way to access our uncertainties on that simulation, including flux, neutrino interaction, and detector uncertainties. The data (and simulated data) has already been pre-selected to contain either a muon candidate or an electron candidate.

Collaboration, MINERvA [Fermi National Accelerator↗

FY 2025 Multidimensional Data Correlation Platform: Unified Software Architecture for Advanced Materials and Manufacturing Technologies Data Management and Processing

The Advanced Materials and Manufacturing Technologies (AMMT) program continues to advance a data-driven approach to demonstrate the utility of additive manufacturing for fabricating components for nuclear applications. A key scientific goal is to leverage data to better understand manufacturing outcomes and thereby improve the performance, reliability, and lifespan of nuclear components. Ultimately, this effort supports the development of standards for certification and qualification of additively manufactured components, enabling broader industry adoption. In support of this objective, the AMMT program is building and deploying a data management platform to record, index, analyze, and make available the manufacturing data generated across the AMMT program. In FY 2023, the team conceptualized the architecture of the platform and, in FY 2024, deployed the first functional version at the Oak Ridge National Laboratory (ORNL) Manufacturing Demonstration Facility (MDF). In FY 2025, the platform was officially opened to all AMMT members. To enable this expansion, core modifications and enhancements were developed, including improvements to the user interface and workflows for data entry and retrieval. Most notably, robust security and access control mechanisms were implemented to protect data and manage information sharing. This effort featured a logging system, protected views, and controlled access mechanisms. This report documents these enhancements and the transition of the platform into program-wide use.

36 MATERIALS SCIENCE↗

Transit Survey - Miami-Dade Metrorail - 2009

This Miami-Dade Metrorail Survey obtained ridership characteristics such as origin-destination patterns, trip purpose, and mode of access and egress. The data obtained from this survey was used to update and validate the Southeast Regional Planning Model (SERPM v6.5) and for transportation planning in the region.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Data Release 1 of the Dark Energy Spectroscopic Instrument

In 2021 May the Dark Energy Spectroscopic Instrument (DESI) collaboration began a 5 yr spectroscopic redshift survey to produce a detailed map of the evolving three-dimensional structure of the Universe between z = 0 and z ≈ 4. DESI’s principal scientific objectives are to place precise constraints on the equation of state of dark energy, the gravitationally driven growth of large-scale structure, and the sum of the neutrino masses, and to explore the observational signatures of primordial inflation. We present DESI DR1, which consists of all data acquired during the first 13 months of the DESI main survey, as well as a uniform reprocessing of the DESI Survey Validation data, which were previously made public in the DESI Early Data Release. The DR1 main survey includes high-confidence redshifts for 18.7M objects, of which 13.1M are spectroscopically classified as galaxies, 1.6M as quasars, and 4M as stars, making DR1 the largest sample of extragalactic redshifts ever assembled. We summarize the DR1 observations, the spectroscopic data-reduction pipeline and data products, large-scale structure catalogs, value-added catalogs, and describe how to access and interact with the data. In addition to fulfilling its core cosmological objectives with unprecedented precision, we expect DR1 to enable a wide range of transformational astrophysical studies and discoveries.

79 ASTRONOMY AND ASTROPHYSICS↗

Predictions for the Detectability of Milky Way Satellite Galaxies and Outer-Halo Star Clusters with the Vera C. Rubin Observatory

We predict the sensitivity of the Vera C. Rubin Observatory Legacy Survey of Space and Time (LSST) to faint, resolved Milky Way satellite galaxies and outer-halo star clusters. We characterize the expected sensitivity using simulated LSST data from the LSST Dark Energy Science Collaboration (DESC) Data Challenge 2 (DC2) accessed and analyzed with the Rubin Science Platform as part of the Rubin Early Science Program. We simulate resolved stellar populations of Milky Way satellite galaxies and outer-halo star clusters over a wide range of sizes, luminosities, and heliocentric distances, which are broadly consistent with expectations for the Milky Way satellite system. We inject simulated stars into the DC2 catalog with realistic photometric uncertainties and star/galaxy separation derived from the DC2 data itself. We assess the probability that each simulated system would be detected by LSST using a conventional isochrone matched-filter technique. We find that assuming perfect star/galaxy separation enables the detection of resolved stellar systems with $M_V$ = 0 mag and $r_{1/2}$ = 10 pc with >50% efficiency out to a heliocentric distance of ~250 kpc. Similar detection efficiency is possible with a simple star/galaxy separation criterion based on measured quantities, although the false positive rate is higher due to leakage of background galaxies into the stellar sample. When assuming perfect star/galaxy classification and a model for the galaxy-halo connection fit to current data, we predict that 89 +/- 20 Milky Way satellite galaxies will be detectable with a simple matched-filter algorithm applied to the LSST wide-fast-deep data set. Different assumptions about the performance of star/galaxy classification efficiency can decrease this estimate by ~7%-25%, which emphasizes the importance of high-quality star/galaxy separation for studies of the Milky Way satellite population with LSST.

79 ASTRONOMY AND ASTROPHYSICS↗

Empowering Geothermal Research: The Geothermal Data Repository's New AI Research Assistant

The Department of Energy's (DOE) Geothermal Data Repository (GDR) team has integrated a Large Language Model (LLM) with the metadata and supporting documents associated with GDR datasets to create an Artificially Intelligent (AI) research assistant. By leveraging work done to make GDR metadata machine-readable and an open-source LLM integration model called the Energy Language Model, developed by the National Renewable Energy Laboratory, AskGDR serves as a virtual research assistant to GDR users. It provides answers to a variety of user-provided questions using natural language processing and generative machine learning. Users can get answers to questions about specific datasets, including inquiries about the equipment, assumptions and methodologies used in the origination of the data; or more abstract questions, such as the applicability of data to specific research fields. AskGDR improves the discoverability of geothermal data by helping guide users to datasets beyond simple keyword searches. It enables users to find data based on properties of the data, discover information contained within supporting documents, and explore data from projects related to their research objectives. This paper will outline the development, integration, output, and efficacy of the AskGDR LLM, including adherence to scientific rigor through improvements designed to increase the accuracy of generated answers, avoid speculation, and provide proper references for all resources used.

access↗

Powering Circularity Through Data Reporting and Collection

Sustainability and Circular Economy have many metrics for evaluation. Calculating mass intensity, energy return on investment, financial payback, and recycling rate for proposed technology changes and lifecycle management can support decision making. Robust data with modeling tools can perform these calculations, informing good decision making. Analyses show that reliability is more critical than recyclability. Improved data gathering and tool accessibility will support our industry to make more circular choices for PV lifecycle management.

14 SOLAR ENERGY↗

SEED: Semantic Energy Exploration and Discovery

The Bioenergy Knowledge Discovery Framework (KDF) hosts a vast repository of specialized data, yet traditional keyword-based search methods often struggle to provide direct answers, requiring significant domain expertise and manual effort to filter through raw documents. To overcome these barriers, this software introduces a semantic search engine that enables both specialists and non-specialists to query the KDF using natural language. By shifting from rigid keyword matching to intent-based retrieval, the tool automatically identifies and ranks the most relevant sources within the database. The system functions by processing natural language queries to extract the most pertinent information, delivering an AI-generated plain-language summary alongside exact supporting quotes from retrieved documents. This integrated approach provides users with immediate, evidence-based answers while eliminating the need for exhaustive manual review. By surfacing direct insights and contextual evidence, the software enhances the usability of existing KDF resources and democratizes access to complex bioenergy data. Ultimately, this semantic search solution accelerates the discovery process and supports faster, more informed decision-making across the bioenergy sector.

Pan, Meiyu (Melrose) [Oak Ridge National Laborator↗

Extending Rucio with modern cloud storage support

Rucio is a software framework designed to facilitate scientific collaborations in efficiently organising, managing, and accessing extensive volumes of data through customizable policies. The framework enables data distribution across globally distributed locations and heterogeneous data centres, integrating various storage and network technologies into a unified federated entity. Rucio offers advanced features like distributed data recovery and adaptive replication, and it exhibits high scalability, modularity, and extensibility. Originally developed to meet the requirements of the high-energy physics experiment ATLAS, Rucio has been continuously expanded to support LHC experiments and diverse scientific communities. Recent R&D projects within these communities have evaluated the integration of both private and commercially-provided cloud storage systems, leading to the development of additional functionalities for seamless integration within Rucio. Furthermore, the underlying systems, FTS and GFAL/Davix, have been extended to cater to specific use cases. This contribution focuses on the technical aspects of this work, particularly the challenges encountered in building a generic interface for self-hosted cloud storage, such as MinIO or CEPH S3 Gateway, and established providers like Google Cloud Storage and Amazon Simple Storage Service. Additionally, the integration of decentralised clouds like SEAL is explored. Key aspects, including authentication and authorisation, direct and remote access, throughput and cost estimation, are highlighted, along with shared experiences in daily operations.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

ComStock Measure Documentation: Thermostat Setbacks During Unoccupied Periods

This report assesses the potential for nationwide adoption of thermostat setbacks in appropriate applications. Building on the 3-year End-Use Load Profiles project to calibrate and validate the U.S. Department of Energy’s ResStock™ and ComStock™ models, this work produces national datasets that enable cities, states, utilities, and other stakeholders to answer a broad range of questions regarding their commercial building stock. ComStock is a highly granular, bottom-up model that uses various data sources, statistical sampling methods, and advanced building energy simulations to estimate the annual sub-hourly energy consumption of the commercial building stock across the United States. The “baseline” model intends to represent the U.S. commercial building stock as it existed in 2018. The methodology of the baseline model is discussed in the ComStock Reference Documentation. The goal of this work is to develop energy efficiency and demand flexibility measures that cover market-ready technologies and study their mass-adoption impact on the baseline building stock. “Measures” refers to various “what-if” scenarios that can be applied to buildings. The results for the baseline and measure scenario simulations are published in public datasets that provide insights into building stock characteristics, operational behaviors, utility bill impacts, and annual and sub-hourly energy usage by fuel type and end use. This report describes the modeling methodology for a single ComStock measure scenario— Thermostat Setbacks During Unoccupied Periods—and briefly introduces key results. The full public dataset can be accessed on the ComStock data lake or via the Data Viewer at comstock.nrel.gov. The public dataset enables users to create custom aggregations of results for their use cases (e.g., filter to a specific county or building type).

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

ComStock Measure Documentation: Fan Static Pressure Reset for Multizone Variable Air Volume Systems

This report assesses the potential for nationwide adoption of a duct static pressure reset in MZ VAV systems in appropriate applications. Building on the 3-year End-Use Load Profiles project to calibrate and validate the U.S. Department of Energy’s ResStock™ and ComStock™ models, this work produces national datasets that enable cities, states, utilities, and other stakeholders to answer a broad range of questions regarding their commercial building stock. ComStock is a highly granular, bottom-up model that uses various data sources, statistical sampling methods, and advanced building energy simulations to estimate the annual sub-hourly energy consumption of the commercial building stock across the United States. The “baseline” model intends to represent the U.S. commercial building stock as it existed in 2018. The methodology of the baseline model is discussed in the ComStock Reference Documentation. The goal of this work is to develop energy efficiency and demand flexibility measures that cover market-ready technologies and study their mass-adoption impact on the baseline building stock. “Measures” refers to various “what-if” scenarios that can be applied to buildings. The results for the baseline and measure scenario simulations are published in public datasets that provide insights into building stock characteristics, operational behaviors, utility bill impacts, and annual and sub-hourly energy usage by fuel type and end use. This report describes the modeling methodology for a single ComStock measure scenario—Fan Static Pressure Reset for Multizone Variable Air Volume (VAV) Systems—and briefly introduces key results. The full public dataset can be accessed on the ComStock data lake or via the Data Viewer at comstock.nrel.gov. The public dataset enables users to create custom aggregations of results for their use case (e.g., filter to a specific county or building type).

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

TSDC: Transportation Secure Data Center: Real-World Data for Planning, Modeling, and Analysis

The Transportation Secure Data Center is a centralized repository for high-resolution transportation data from hundreds of travel and transit surveys and studies. It makes vital transportation data broadly available to users while preserving the privacy of survey participants. It houses surveys and studies conducted by state departments of transportation, metropolitan planning organizations, transit agencies, cities, and other public agencies. Meanwhile, the Livewire Data Platform empowers research, industry, and academic partners to easily and securely preserve, maintain, share, discover, and gain access to transportation and mobility data. Livewire accommodates a range of datasets, including behavioral, experimental, model, analytical, and raw data at the vehicle, traveler, and system levels. Datasets support mobility research and planning spanning urban science, connected and automated vehicles, fueling and charging infrastructure, mobility decision science, multimodal transportation, vehicle efficiency, and more.

33 ADVANCED PROPULSION SYSTEMS↗

Qualitative Risk Assessment of Legacy Wells within the Estimated Prairie State Generating Company Area of Review

This report details the digitization of a legacy wellbore database, including data processing assumptions, parameter estimation, and risk assessment methodology. The database, comprising 6,454 documents, was provided by ISGS. It includes valuable data from the Prairie State Generating Company (PSGC) and One Earth Energy (OEE) sites of the CarbonSAFE Phase III – Illinois Storage Corridor project. The report focuses on wells within a 15-mile radius from the Lively Grove #1 (LG#1) well at PSGC site, evaluating subsurface conditions and potential risks. A total of 4,386 wellbores within 15 miles of the LG#1 well were filtered based on depth and formation codes. LG#1 is the stratigraphic well at the PSGC site drilled in 2021. Ninety-four (94) wells penetrating the Maquoketa Shale Group (the primary confining unit) within the estimated area-of-review (AoR) for the PSGC site were evaluated using a qualitative risk assessment (QRA) methodology. The QRA developed by Arbad et al. 2022 focuses on legacy wells within the AoR and categorizes them based on well construction details. The QRA identifies wells that need immediate attention by categorizing them based on penetration depth and protection. Wells within the AoR were categorized into nine groups based on penetrations and protections. These categories range from Type 1 wells, with no documentation, to Type 9 wells, which do not penetrate the primary confining unit or storage reservoir (unit). Well accessibility within the AoR varies based on well status, including Dry & Abandoned (DA), Plugged & Abandoned (PA), Injection (INJ), Oil/Gas Producing (PROD), and Observation (Obs) wells. Accessibility levels were determined by well construction, with DA wells being the least accessible and Observation wells the most accessible, impacting gas leakage detection possibilities. Remedial action priority of wells decreases from Type 1 to Type 9 wells. Type 1 to Type 6 wells with status DA and PA require immediate attention, while Type 7 and Type 8 wells are low priority. A risk matrix used to prioritize corrective actions for legacy wells is proposed to categorize wells within an AoR based on penetrations, protections, and accessibility. The methodology involves data acquisition, well categorization into nine types, and determining CO 2 leakage pathways using well schematics and geospatial mapping. This approach is particularly useful for managing the integrity of legacy wells throughout the lifecycle of a Carbon Capture and Storage (CCS) project. A qualitative risk assessment of 94 wells within the AoR of the PSGC site identified 54 wells with high priority for corrective action due to penetration of the primary containment seal. The assessment utilizes color-coded maps to categorize well types and prioritize corrective actions, providing a comprehensive analysis. Schematics of wells penetrating the primary confining unit were drawn, and leakage pathways were identified. Details of all wells penetrating the confining zone are provided in the appendix, including information on well types, plugging, and casing status.

01 COAL, LIGNITE, AND PEAT↗

Establishing Data Analysis Pipeline for Bulk ATAC-Seq Datasets

We developed an analysis pipeline for transposase-accessible chromatin sequencing (ATAC-Seq) data derived from bulk samples, which brings together publicly available R packages in addition to command-line tools designed for analysis of bulk ATAC-Seq data and can be run on any computer running a Linux-like operating system such as Ubuntu or Apple OSX.

97 MATHEMATICS AND COMPUTING↗

The genomic footprints of wild Saccharum species trace domestication, diversification, and modern breeding of sugarcane

Sugarcane is a major crop of unclear origins due to its complex polyploid interspecific genome. We analyzed genome ancestries using whole-genome sequence data from 390 representative accessions based on repeated k-mers and chloroplast phylogeny. The results provided evidence that Saccharum officinarum was domesticated in the New Guinea region from the S. robustum wild species and revealed that its genome is a mosaic involving different S. robustum subgroups. We discovered a wild Saccharum contributor to most modern cultivars, likely originating from East Melanesia. We highlighted two early centers of sugarcane diversification associated with human transport, one in continental Asia through hybridization with different S. spontaneum subgroups and one in the Melanesian and Polynesian islands via hybridization with the discovered ancestor and Miscanthus. Finally, we revealed the genome ancestry of modern cultivars, highlighting untapped wild Saccharum diversity as a source of alleles for breeding programs.

Garsmeur, Olivier [CIRAD, Montpellier (France). Ag↗

Has Reducing Ship Emissions Brought Forward Global Warming?

Abstract Ships brighten low marine clouds from emissions of sulfur and aerosols, resulting in visible “ship tracks”. In 2020, new shipping regulations mandated an ∼80% reduction in the allowed fuel sulfur content. Recent observations indicate that visible ship tracks have decreased. Model simulations indicate that since 2020 shipping regulations have induced a net radiative forcing of +0.12 Wm −2 . Analysis of recent temperature anomalies indicates Northern Hemisphere surface temperature anomalies in 2022–2023 are correlated with observed cloud radiative forcing and the cloud radiative forcing is spatially correlated with the simulated radiative forcing from the 2020 shipping emission changes. Shipping emissions changes could be accelerating global warming. To better constrain these estimates, better access to ship position data and understanding of ship aerosol emissions are needed. Understanding the risks and benefits of emissions reductions and the difficultly in robust attribution highlights the large uncertainty in attributing proposed deliberate climate intervention.

54 ENVIRONMENTAL SCIENCES↗

Unified architecture for quantum lookup tables

Quantum access to arbitrary classical data encoded in unitary black-box oracles underlies interesting data-intensive quantum algorithms, such as machine learning or electronic structure simulation. The feasibility of these applications depends crucially on gate-efficient implementations of these oracles, which are commonly some reversible versions of the Boolean circuit for a classical lookup table. Here, we present a general parametrized architecture for quantum circuits implementing a lookup table that encompasses all prior work in realizing a continuum of optimal trade-offs between qubits, non-Clifford gates, and error resilience, up to logarithmic factors. Our architecture assumes only local 2D connectivity, yet recovers results, with the appropriate parameters, polylogarithmic error scaling. We also identify regimes, such as simultaneous sublinear scaling, in all parameters. These results enable tailoring implementations of the commonly used lookup table primitive to any given quantum device with constrained resources.

quantum circuits↗

Material Needs and Measurement Challenges for Advanced Semiconductor Packaging: Understanding the Soft Side of Science

This Perspective builds upon insights from the National Institute of Standards and Technology (NIST)-organized workshop, “Materials and Metrology Needs for Advanced Semiconductor Packaging Strategies,” held at the 35th annual Electronics Packaging Symposium in Binghamton, NY, on September 5, 2024. It outlines critical challenges and opportunities related to polymer-based “soft” materials in advanced semiconductor packaging, with emphasis on polymer science, measurement science (metrology), and the strategic development of Research-Grade Test Materials (RGTMs). These efforts, led by the NIST CHIPS team, aim to advance the fundamental understanding of structure-property-processing relationships, promote standardized guidelines and innovative methods for material characterization, and accelerate the development, qualification, and adoption of next-generation packaging materials. The Perspective also distills key insights from the panel discussion with industry experts, emphasizing the need for close collaboration among materials scientists, process engineers, and metrology experts to enable a holistic strategy, further highlighting the importance of cross-sector partnerships among industry, academia, and government to address pressing challenges in packaging materials and processes.

97 MATHEMATICS AND COMPUTING↗