Search NASA⌕ Search

SEARCH · Search NASA

Results for “Data Collection”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 343 records · Page 19

A Decision Support System to Compile Environmental Mitigations from Hydropower Licensing Documents

The process of deciphering, extracting, and compiling information from texts dense with domain-specific terminology and technical jargon is a challenging endeavor. It demands considerable expertise and deep knowledge in the respective field, resulting in a labor-intensive process when executed by humans. Furthermore, the task of identifying multiple class labels in extensive texts presents a challenge due to intra- and inter-reader variability, making the process time-consuming and costly.We’re introducing a user-friendly graphical interface, fortified with a BERT model-powered decision support system. This advanced system aims to augment efficiency, curtail data collection time, and sustain high precision in data acquisition. It is instrumental in deciphering and synthesizing intricate texts teeming with a spectrum of expressions, even within similar mitigation categories. Such tasks traditionally demand substantial human effort and specialized knowledge in the domain.Our system is specifically engineered for the task of extracting environmental mitigation information to promote sustainable hydropower development from licenses issued by the Federal Energy Regulatory Commission (FERC). These license documents are comprehensive, each containing over 15,000 words and requiring the identification of 135 different class labels. We anticipate that our system will boost reading speed, improve the consistency of classification outputs among readers, and contribute to the development of a robust scientific database of environmental mitigations associated with the 2,000+ non-federal hydropower facilities licensed by FERC in the United States.

Yoon, Hong-Jun [ORNL] (ORCID:0000000254505878)↗

Scalable Multi-Facility Workflows for Artificial Intelligence Applications in Climate Research

Earth observation satellites and earth system models are sources of vast, multi-modal datasets that are invaluable for advancing climate and environmental research. However, their scale and complexity pose significant challenges for processing and analysis. In this paper we discuss our experiences in developing and using a scientific research application using an automated multi-facility workflow that orchestrates data collection, preprocessing, artificial intelligence (AI) inferencing, and data movement across diverse computational resources, leveraging the Advanced Computing Ecosystem Testbed at the Oak Ridge Leadership Computing Facility (OLCF). We demonstrate that our workflow can be seamlessly integrated and orchestrated across research facilities managed by different federal agencies, thus allowing users to extract new scientific insights from climate datasets. The experimental results indicate that the multi-facility workflow significantly reduces processing time, enhances scalability, and maintains high efficiency across varying workloads. Notably, our workflow processes 12,000 high-resolution satellite images in just 44 seconds using 80 workers distributed across 10 nodes on the OLCF systems. Such high throughput is essential for dynamic tokenization and sharding of petascale satellite data for distributed AI model training and inferencing at scale across thousands of GPUs.

Kurihana, Takuya [ORNL] (ORCID:0000000156698565)↗

Automating Log Synthesis and Visualization with Python and Splunk

The goal of this project is to automate log analysis by utilizing Splunk, Bash, and Python together. Simplifying the monitoring and analysis of network traffic was the main goal. In order to accomplish this, a Bash script was created to use 'tcpdump' to automate network sniffing. It also included a 24-hour file rotation mechanism to effectively manage the pcap files that were generated. After that, a Python script was written to read these pcap files and retrieve pertinent data about network traffic. After processing the collected data, Splunk is used to summarize the important metrics and visualize said information with relevant graphs.

99 GENERAL AND MISCELLANEOUS↗

Smart Contracts for Power Grid Applications Using the Advanced DLT Cyber Grid Guard Testbed

In this study is presented two power system applications with distributed ledger technology (DLT) and smart contracts (SC) that were assessed in a Cyber-Grid-Guard System (CGGS) advanced testbed, with protective relays, power meters, communication devices, DLT devices, synchronized time source, clock displays and real time simulator. This CGGS testbed was set in the Advanced Protection Lab, 252 lab space of the Grid Research Integration and Deployment Center (GRID- C), at Oak Ridge National Laboratory. In power grids, customer-owned distributed energy resources (DERs) are more frequent than in the past, and the numbers of points of interconnection (POI) with customer-owned DERs have increased. Disruptive operation from DERs presents a risk to grid operations, and protective relays located at the POI are used to isolate out-of-tolerance or poorly behaving of DERs. Ensuring the integrity of data from the relays at the POI, and DLT could enhance the security of the power grids. The first application is a SC to define and control the allowable total power factor (TPF) of the DER (wind farm) output, and the terms of the SC are implemented using DLT with a CGGS for a customer-owned DER. The TPF SC was implemented by the CGGS using DLT. The experimental model was performed with a real-time simulator using a CGGS and relay in-the-loop. The data collected from the CGGS were used to execute the TPF SC. The TPF limits were between +0.9 and +1.0, and the breakers’ operation in the POI was controlled by the relay using the SC. The events were collected from the real-time simulator, CGGS, and SEL 700GT relay to validate a successful application of the TPF SC using DLT. The second application is a SC to measure and control the allowable voltage service limits (VSL) by the CGGS using DLT. The tests were performed by using a real-time simulator, CGGS and relay in-the-loop. The data was collected from the CGGS that executed the SC. The main constraints were defined based on ANSI C84.1 service voltage limits, and the operation of the breakers in the POI. The events were collected from the CGGS, and SEL 700GT relay to assess a successful operation of the VSL SC using DLT.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Building a framework to genetically characterize “feather spots” and understand demographic impacts of solar energy sites on migratory bird populations

The lack of data on the impact of utility-scale solar facilities on avian species and populations adds to the cost of siting and operation. As much as 32 percent of the avian biological material (feathers and carcasses) recovered from solar facilities remain unidentified, because they often take the form of “feather spots”. Feather spots are remains of impacted animals that can be separated into two broad categories: 1) those remains that may be visually identified to a species, or 2) those that cannot be visually identified to a species due to degradation from the environment and/or scavenger activity (listed as “unknown”). Even when feather spots can be identified to species, they cannot be visually assigned to particular breeding populations. In some cases, it is unknown whether multiple feather spots represent single or multiple individuals. This project’s objectives were to: 1. Use a developed, genetic-based technique to identify and determine the species, population of origin, and number of individuals found in feather spots recovered from solar facilities. 2. Implement collected data and resulting analyses to develop a publicly accessible web-based decision-making tool that can be used by the solar industry, regulators and other stakeholders to inform siting, mitigation, and conservation management efforts. 3. Establish a not-for-profit fee-for-service center at UCLA to ensure collection and identification of feather spots continue after the project period of performance. During the Project Period, we proposed to establish a pipeline for collecting, transporting, and storing of avian biological material collected at solar facilities and the collection and identification of feather spots to species and individual. We proposed the development of a genetic-based framework that would recover viable DNA from feather spots, amplify this DNA (i.e., make millions of copies of the original DNA), and use it to match the resulting sequences to a national database of known species of birds. The result would be the identification of feathers spots that were previously unidentified, and the incorporation of these samples into a larger database that included all samples recovered from solar facilities. The resulting report (below) details the result of this work and its alignment with proposed activities. We proposed the use of the data collected to assess the comparative risk to specific species or populations of species from solar facilities. For some species, we have already identified genomic markers of specific breeding populations and developed “genoscapes,” maps of unique genetic variation across the full breeding range of a species. We used these (previously and newly developed) genoscapes to probabilistically link a feather spot to the specific breeding populations from which it originated (assignment probabilities range from 75%-100% depending on species and population groups). For those species without genoscapes, we developed a vulnerability and susceptibility estimate that determines the relative local and regional risk to populations that are in geographic proximity to solar facilities, using citizen science data (Breeding Bird Survey (BBS) and eBird). These two feather spot processing pipelines (see Figure 1 below) provide quantitative estimates as to the numbers of individuals from a given population of origin that are affected by solar facilities, and ultimately can reduce costs to the consumer by reducing the industry costs associated with mitigation and siting strategies for future solar energy development.

14 SOLAR ENERGY↗

BRE‐X Emissions Database for End‐of‐Life Scenarios of Selective Building Construction Materials to Enable Circular Economy in Construction

In the United States, construction and demolition debris predominately end up in landfills with minimal end‐of‐life Re‐X (recover, recycle, reuse, etc.) scenarios, resulting in large environmental impacts and lost opportunities for material recovery. Except for concrete and metals, which seem to have a few well‐defined end‐of‐life pathways, there seems to be a lack of well‐documented end‐of‐life scenarios for other construction materials, let alone their emissions data. Hence, there is a need for documented end‐of‐life Re‐X scenarios and end‐of‐life data of more building materials to motivate widespread use of Re‐X strategies in building design. This paper outlines the efforts of the National Renewable Energy Laboratory, Carbon Leadership Forum, Building Transparency, and Skidmore, Owings & Merrill to (a) create an open‐access BRE‐X (Building Re‐X) end‐of‐life emissions database consisting of greenhouse gas emissions data associated with various end‐of‐life scenarios for a select list of high‐impact building construction materials, and (b) integrate the BRE‐X end‐of‐life emissions database with CAD/BIM/LCA tools for evaluating various end‐of‐life scenarios. The paper also presents a few existing life cycle inventory databases that contain sparse amounts of end‐of‐life data for a few construction materials and their limitations in terms of scaling and data consolidation. Finally, a sample of how the collected data can be ingested into whole‐building LCA tools using open data formats and a public access link to the BRE‐X end‐of‐life emissions database is also included.

36 MATERIALS SCIENCE↗

Advancing Asset Management in Water Infrastructure Systems

Aging water system infrastructure, including drinking water, wastewater, and stormwater, poses a growing challenge for utilities and municipalities. These water systems have well documented challenges with respect to their age, condition, and level of service. ASCE annual report cards consistently rate these infrastructure systems in the United States as underfunded, overcapacity, or past service life (ASCE 2025 Report Card). For example, Chini and Stillwell (2017) estimated that the mean water loss in drinking water systems, i.e., non-revenue water, is approximately 16% across the United States. These concerns are not just relegated to the United States, with Courtenay, British Columbia, identifying 17% of their water main pipes as in a ‘poor’ condition state, defined as a category condition 5 out of 5 (City of Courtenay, 2024). These cases illustrate the challenges utilities are facing to manage extensive networks of infrastructure to deliver a consistent and high level of service. For buried infrastructure such as water systems, studies suggest that preventative interventions can lead to lower maintenance costs and fewer service disruptions (Mazumder et al, 2018; Li et al, 2014). The demonstrated need and benefit of appropriately applied asset management is juxtaposed against the relatively sparse literature that evaluates water systems within an asset management construct. Since 2020, just 37 papers specifically reference asset management in the Journal of Water Resources Planning and Management. Of those, only a few specifically look to develop strategies for improved asset management. Therefore, we highlight four key research areas that represent opportunities for advancement of asset management research for water systems. First, advances in condition assessment and forecasting are needed to better estimate asset deterioration using diverse datasets. Second, machine learning (ML) and artificial intelligence (AI) hold promise for predictive maintenance and investment prioritization, though questions of generalizability and model transparency remain. Third, applying a value of information framework can guide utilities in making cost-effective sensor deployment and data collection decisions, to direct monitoring strategies towards data-informed asset management decisions. Finally, integrated infrastructure management is critical, requiring coordinated planning with other infrastructure systems and stakeholder engagement to reduce costs and enhance service delivery.

Chini, Christopher M.↗

FY 2024 Multidimensional Data Correlation Platform Data Management Infrastructure Progress: Materials Laboratory

This report provides an inventory of the equipment available at the ORNL Manufacturing Demonstration Facility (MDF) for sample preparation and material characterization, including both destructive and non-destructive techniques that generate critical data to support the development of the Multi-Dimensional Data Correlation (MDDC) framework. The success of the MDDC framework depends heavily on the quality and completeness of the data it can access. Therefore, it is essential to establish a comprehensive inventory of the technologies available to the Advanced Materials and Manufacturing Technologies (AMMT) multi-laboratory team. This starts by gathering information about the types of data they produce, the data collection and transfer protocols used, file formats, and data storage requirements for experiments. This information is then carefully evaluated to create the operations and trackables elements of the Damara Tern platform, which is the foundation of the MDDC framework.

36 MATERIALS SCIENCE↗

An Overview of the Molten Salt Thermal Properties Database–Thermophysical, Version 4.0 (MSTDB-TP V.4.0)

A central repository of thermophysical and thermochemical properties of molten salt compositions of relevance to molten salt reactors (MSRs) is vital in supporting the broad community of MSR developers, who are at various stages of developing and deploying their reactor designs. In general, these MSR designs differ significantly from developer to developer (e.g., with respect to the hardness of the neutron spectra, level of fissile loading, target multicomponent temperatures and power levels, and moderating capabilities). Therefore, the fuel and coolant salts being considered vary greatly: they may be chlorides or fluorides, they utilize different actinides at different ratios, and the cations in the melt are selected based on perceived advantages and disadvantages. Considering the general need for thermal properties, and the vastness of the array of potential candidate salt mixtures, the Molten Salt Thermal Properties Database (MSTDB) was initiated in 2018 with the goal of providing thermophysical and thermochemical characterization of key molten salt compounds and mixtures across their temperature and compositional domains. The MSTDB is thus divided into the thermophysical arm (MSTDB-TP) and the thermochemical arm (MSTDB-TC). The MSTDB is an effort funded by the Department of Energy, Office of Nuclear Energy (DOE-NE) Nuclear Energy Advanced Modeling and Simulation (NEAMS) program, and the MSR Campaign. This report provides an overview of the MSTDB-TP v4.0 in terms of the data contained within, the state of the tools used to access the data, the availability of predictive models that leverage the raw data in the database, the preliminary status of developmental efforts that are currently underway, and an account of future goals for MSTDB-TP. The primary goal for the update from MSTDB-TP v.3.1 to v4.0 was the incorporation of surface tension data into the database; this property is important for thermal hydraulics modeling and species transport in other tools that have been developed under the NEAMS program. A breakdown of the surface tension data that have been added into MSTDB-TP v4.0 is provided herein, and the manner in which the quality of the data has been assessed is also documented. For MSTDB-TP v4.0, newly published thermophysical property data—primarily from collaborative experimental efforts under the MSR Campaign—have been incorporated into the database, and the resulting expansion is documented here. Because of the size to which MSTDB-TP has grown, the raw data format has now been recast into JavaScript Object Notation (JSON) format for easier connection with the MSTDB-TP application programming interface (API). Saline; the pre-existing comma-separated value (CSV) format has been deprecated but is still maintained, accessible, and up to date. As a final effort in packaging the MSTDB-TP v4.0 update, the graphical user interface (GUI) for MSTDB has been updated to allow full accessibility to the density and viscosity predictive models, which are based on Redlich-Kister expansions of MSTDB-TP raw data. Some other major aspects of this report, in terms of preliminary and future work, include: (1) documentation of the formalism and preliminary testing of a kinetic theory model that may act as a predictive model for thermal conductivity; (2) documentation of the candidate predictive models that may be considered in the future for surface tension, making use of the surface tension data now in MSTDB-TP v4.0; (3) a preliminary account of a data collection process that will enable the filling of additional gaps within MSTDB-TP, namely with data which have been collected computationally (e.g., through ab initio molecular dynamics).

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Uinta Basin CarbonSAFE II: Storage Complex Feasibility (Final Report)

The primary objective of this CarbonSAFE Phase II project was to establish the technical and commercial feasibility of a commercial-scale CO 2 geological storage complex for Deseret Power Electric Cooperative Bonanza Power Plant and other CO 2 sources in the northeast Uinta Basin, Utah, with the goal to securely store at least 50 million metric tons of captured CO 2 and accelerate CO 2 capture, utilization, and storage (CCUS) deployment. The project team established high-potential technical and commercial feasibility for a storage site within the east Uinta Basin (Utah), in the Cretaceous sandstones (Frontier, Dakota, and Buckhorn), Entrada Sandstone, Nugget Sandstone, and/or Weber Sandstone southwest of the Bonanza coal-fired power plant. This project collected and analyzed state-of-the-art data to characterize the storage complex consistent with Environmental Protection Agency (EPA) permitting standards. The team conducted extensive analog studies, outcrop mapping, and data sampling, which largely contributed to understanding the subsurface lithology and facies. Existing data were obtained and assessed from Utah Division of Oil, Gas, and Mining (DOGM), Utah Geological Survey (UGS), Colorado Geological Survey (CGS), U.S. Geological Survey (USGS), and EPA. These data were analyzed using state-of-the-art CCUS technologies for Societal Considerations, Site Characterization, Modeling and Simulations, Risk Assessment, Management and Monitoring, potential Underground Injection Control (UIC) Class VI Well Permitting, and Technical/Economic Feasibility. Through these high-resolution data collection and feasibility studies, this project was expected to provide a reference for initiating Underground Injection Control (UIC) and other commercial-scale geological storage permitting processes in the Western United States, ultimately contributing to the nation's decarbonization goals through low-risk, cost-effective commercial-scale carbon capture, utilization, and storage (CCUS) projects.

42 ENGINEERING↗

Expanding NSI searches at NOvA

NOvA is an accelerator-based long-baseline neutrino experiment with two functionally equivalent detectors, designed to study neutrino oscillations. NOvA has also been able to look for signals of new physics like non-standard interactions with matter, setting constraints on the parameters governing that beyond-standard neutrino-physics phenomenon. With data collection progressing, and an upgraded analysis including new data samples and improved understanding of the systematics, we are able to further advance our quest of constraining new physics. Here we will present an update on the analysis status of NOvA on the NSI parameters when the addition of more data and a set of low-energy electron neutrino events not considered in our previous analysis.

Acero Ortega, Mario Andres [U. Atlantico, Barranqu↗

EMPHATIC Silicon Strip Detector Efficiencies

EMPHATIC is an experiment at Fermilab which aims to reduce current neutrino flux uncertainties. This report discusses the limitations current neutrino flux uncertainties places on large scale neutrino experiments, provides background on the EMPHATIC experiment, and details the project of determining the efficiency of the Silicon Strip Detectors (SSDs) used in EMPHATIC. As part of the data analysis process and in order to increase the accuracy of EMPHATIC’s simulations a representation of efficiency of each SSD is required. To achieve this a data-driven analysis was performed on EMPHATIC's collected data using the Root and Art frameworks. Visual and numerical representations of efficiency were determined. The average efficiency over all SSDs is 98.58\%, however this number deflated as it includes known bad channels.

Olson, Virginia [Illinois U., Urbana (main)]↗

Urban Parameters Arizona Urban Corridor 100m

132 Urban parameters based on building physical dimensions and location were generated for the cities in six Arizona Counties at 100m resolution using the NATURF model. To use the binary file with WRF, the binary file and the index file must be placed in their own directory in WRF_GEOG and accessed in the same way NUDAPT44 would be accessed.

Dumas, Melissa [ORNL] (ORCID:0000000233190846)↗

Urban Parameters Los Angeles County 100m version 2

132 Urban parameters based on building physical dimensions and location were generated for the city of Los Angeles at 100m resolution using the NATURF model. To use the binary file with WRF, the binary file and the index file must be placed in their own directory in WRF_GEOG and accessed in the same way NUDAPT44 would be accessed. The kmz file can be visualized on Google Earth.

Sweet-Breu, Levi [Baylor University]↗

Development of Data Reporting Standards for High-Temperature Gas-cooled Reactor (HTGR) Nuclear Energy University Program (NEUP) Thermal-Fluid Experiments

Since 2009, the U.S. Department of Energy (DOE) Office of Nuclear Energy's Nuclear Energy University Program (NEUP) has been at the forefront of nuclear research, specifically concentrating on advancing high-temperature gas-cooled reactor (HTGR) technologies. By Fiscal Year 2023, NEUP has authorized 35 projects dedicated to HTGR research, each contributing significantly to the enhancement of our understanding of this technology. The outcomes of these diverse projects have been disseminated through final NEUP reports, peer-reviewed journal articles, and presentations at academic conferences, forming a comprehensive tapestry of knowledge. Despite the substantial value of these findings, their dissemination has been fragmented, posing challenges for accessibility to researchers and policymakers and leading to underutilization of DOE investments. Recognizing this critical gap and its potential consequences for the future of nuclear research, the Advanced Reactor Technologies (ART) Gas-Cooled Reactor (GCR) program conducted an extensive survey of completed and ongoing HTGR NEUP projects. This survey enabled the compilation of crucial data, resulting in the development of a specialized public-access database tailored for computational fluid dynamics and system code validation, specifically designed for HTGR applications. However, the data collection process revealed a significant challenge in central data organization due to individual researchers from different institutes employing varying logics and preferences for recording and documenting experimental data. Consequently, an urgent need has been identified to establish a standardized reporting format for HTGR experimental projects. Addressing this issue is essential for enhancing collaboration, maximizing the impact of DOE investments, and ensuring the seamless advancement of HTGR technologies in nuclear research.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

HTGR Validation: NEUP Survey and Database - Data Reporting Standard for HTGR Thermal-Fluid Experiments

Since 2009, the U.S. Department of Energy (DOE) Office of Nuclear Energy's Nuclear Energy University Program (NEUP) has been at the forefront of nuclear research, specifically concentrating on advancing high-temperature gas-cooled reactor (HTGR) technologies. By Fiscal Year 2024, NEUP has authorized 36 projects dedicated to HTGR research, each contributing significantly to the enhancement of our understanding of this technology. The outcomes of these diverse projects have been disseminated through final NEUP reports, peer-reviewed journal articles, and presentations at academic conferences, forming a comprehensive tapestry of knowledge. Despite the substantial value of these findings, their dissemination has been fragmented, posing challenges for accessibility to researchers and policymakers and leading to underutilization of DOE investments. Recognizing this critical gap and its potential consequences for the future of nuclear research, the Advanced Reactor Technologies (ART) Gas-Cooled Reactor (GCR) program conducted an extensive survey of completed and ongoing HTGR NEUP projects. This survey enabled the compilation of crucial data, resulting in the development of a specialized public-access database tailored for computational fluid dynamics and system code validation, specifically designed for HTGR applications. However, the data collection process revealed a significant challenge in central data organization due to individual researchers from different institutes employing varying logics and preferences for recording and documenting experimental data. Consequently, an urgent need has been identified to establish a standardized reporting format for HTGR experimental projects. Addressing this issue is essential for enhancing collaboration, maximizing the impact of DOE investments, and ensuring the seamless advancement of HTGR technologies in nuclear research.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Determining the Efficiency of EMPHATICs Silicon Strip Detectors (SSDs)

EMPHATIC is an experiment at Fermilab which aims to reduce current neutrino flux uncertainties. This report discusses the limitations current neutrino flux uncertainties places on large scale neutrino experiments, provides background on the EMPHATIC experiment, and details the project of determining the efficiency of the Silicon Strip Detectors (SSDs) used in EMPHATIC. As part of the data analysis process and in order to increase the accuracy of EMPHATIC’s simulations a representation of efficiency of each SSD is required. To achieve this a data-driven analysis was performed on EMPHATIC's collected data using the Root and Art frameworks. Visual and numerical representations of efficiency were determined. The average efficiency over all SSDs is 98.58\%, however this number deflated as it includes known bad channels.

Olson, V. [Illinois U., Urbana (main)]↗

CatTestHub: A benchmarking database of experimental heterogeneous catalysis for evaluating advanced materials

The ability to quantitatively compare newly evolving catalytic materials and technologies is hindered by the widespread availability of catalytic data collected in a consistent manner. While certain catalytic chemistries have been widely studied across decades of scientific research, quantitative comparisons based on literature information is hindered by variability in reaction conditions, types of reported data, and reporting procedures. Here, we present CatTestHub, an open-access database dedicated to benchmarking experimental heterogeneous catalysis data. Combining systematically reported catalytic activity data for selected probe chemistries, with relevant material characterization and reactor configuration information, the database provides a collection of catalytic benchmarks for distinct classes of active site functionality. Through key choices in data access, availability, and traceability, CatTestHub seeks to balance the fundamental information needs of chemical catalysis and the FAIR data design principles. Details of the database architecture and the means through which to navigate it are presented, highlighting examples of catalytic insights readily drawn from the available benchmarking data. In its current iteration, CatTestHub spans over 250 unique experimental data points, collected over 24 solid catalysts, that facilitated the turnover of 3 distinct catalytic chemistries. Here, a roadmap is presented through which to expand the open-access platform that serves as a community wide benchmark, primarily through continuous addition of kinetic information on select catalytic systems by members of the heterogeneous catalysis community at large.

Benchmark↗