Search NASA⌕ Search

SEARCH · Search NASA

Results for “data archive”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 145 records · Page 8

Cultural Shifts in High Energy Physics Collaboration from the Cold War to the Present: A Historical and Philosophical Perspective

Here, this article employs empirical history and the philosophy of science to study cultural convergences and divergences in international collaborations in high energy physics. We examine two cases: (1) E-36, an experiment on small angle proton-proton scattering conducted during the Cold War at the National Accelerator Laboratory (NAL) in the USA by Soviet and US scientists and (2) an ongoing collaborative experiment, NICA, at the Joint Institute for Nuclear Research (JINR, Dubna), which is a project devoted to heavy-ion physics. The JINR, particularly its Laboratory of High Energy Physics (formerly the “Laboratory of High Energies”) is the main mediating actor between these two cases (i.e., E-36 and NICA), as the majority of Soviet participants in E-36 were representatives of the Institute. Using empirical data collected through archival searches, field observations conducted at JINR in 2018–2019, and in-depth interviews, we tell a story of cultural differences in high energy physics by applying the concepts of ‘trading zones’ (P. Galison) and the translation of interests in actor-networks (B. Latour, M. Callon and others). We analyze three types of cultural diversity (specialization, nationality, and generational) in light of the implications of temporal context and the dichotomy between East and West, showing the roles cultural diversity plays in scientific collaboration (which is an integral part of as well as obstacle to scientific research that can nevertheless provide learning opportunities). Our study aims to demonstrate how disunity and diversity may function in scientific research and how high energy physics collaborations can remain productive despite sometimes deep divergences, including those between East and West.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Outcomes and Conclusions from the 2022 AM Bench Measurements, Challenge Problems, Modeling Submissions, and Conference

The Additive Manufacturing Benchmark Test Series (AM Bench) provides rigorous measurement data for validating additive manufacturing (AM) simulations for a broad range of AM technologies and material systems. AM Bench includes extensive in situ and ex situ measurements, simulation challenges for the AM modeling community, and a corresponding conference series. In 2022, the second round of AM Bench measurements, challenge problems, and conference were completed, focusing primarily upon laser powder bed fusion (LPBF) processing of metals, and both material extrusion processing and vat photopolymerization of polymers. In all, more than 100 people from 10 National Institute of Standards and Technology (NIST) divisions and 21 additional organizations were directly involved in the AM Bench 2022 measurements, data management, and conference organization. The international AM community submitted 138 sets of blind modeling simulations for comparison with the in situ and ex situ measurements, up from 46 submissions for the first round of AM Bench in 2018. Analysis of these submissions provides valuable insight into current AM modeling capabilities. The AM Bench data are permanently archived and freely accessible online. The AM Bench conference also hosted an embedded workshop on qualification and certification of AM materials and components.

36 MATERIALS SCIENCE↗

The LED calibration systems for the mDOM and D-Egg sensor modules of the IceCube Upgrade: Design, production, testing and use in module calibration

The IceCube Neutrino Observatory, instrumenting about 1 km 3 of deep, glacial ice at the geographic South Pole, is due to be enhanced with the IceCube Upgrade. The IceCube Upgrade, to be deployed during the 2025/26 Antarctic summer season, will consist of seven new strings of photosensors, densely embedded near the bottom center of the existing array. Aside from a world-leading sensitivity to neutrino oscillations, a primary goal is the improvement of the calibration of the optical properties of the instrumented ice. This calibration will be applied to the entire archive of IceCube data, improving the angular and energy resolution of the detected neutrino events. For this purpose, the Upgrade strings include a host of new calibration devices. Aside from dedicated calibration modules, several thousand LED flashers have been incorporated into the photosensor modules. We describe the design, production, and testing of these LED flashers before their integration into the sensor modules as well as the use of the LED flashers during lab testing of assembled sensor modules.

Analogue electronic circuits↗

AT 2018dyk: tidal disruption event or active galactic nucleus? Follow-up observations of an extreme coronal line emitter with the Dark Energy Spectroscopic Instrument

We present fresh insights into the nature of the tidal disruption event (TDE) candidate AT 2018dyk. AT 2018dyk has sparked a debate in the literature around its classification as either a bona-fide TDE or as an active galactic nucleus (AGN) turn-on state change. A new follow-up spectrum taken with the Dark Energy Spectroscopic Instrument, in combination with host-galaxy analysis using archival SDSS–MaNGA data, supports the identification of AT 2018dyk as a TDE. Specifically, we classify this object as a TDE that occurred within a gas-rich environment, which was responsible for both its mid-infrared (MIR) outburst and development of Fe coronal emission lines. Comparison with the known sample of TDE-linked extreme coronal line emitters (TDE-ECLEs) and other TDEs displaying coronal emission lines (CrL-TDEs) reveals similar characteristics and shared properties. For example, the MIR properties of both groups appear to form a continuum with links to the content and density of the material in their local environments. This includes evidence for a MIR colour–luminosity relationship in TDEs occurring within such gas-rich environments, with those with larger MIR outbursts also exhibiting redder peaks.

79 ASTRONOMY AND ASTROPHYSICS↗

High fidelity blade-resolved and actuator line data from a 16 turbine wind farm simulation using ExaWind

This data was generated with the ExaWind code suite (https://github.com/Exawind) as a demonstration of a large, 16 turbine wind farm simulation, calculated using two different levels of fidelity. The lower level of fidelity approach uses an actuator line approach to represent the turbines, and was simulated with AMR-Wind (https://github.com/Exawind/amr-wind/) as the background flow solver, coupled to OpenFAST (https://github.com/OpenFAST/openfast). The higher level of fidelity simulation uses a blade-resolved approach, and is done using AMR-Wind, Nalu-Wind (https://github.com/Exawind/nalu-wind), OpenFAST, and TIOGA (https://github.com/Exawind/tioga). In the blade-resolved simulation, ExaWind couples together a background flow solver, AMR-Wind, and a near-body solver, Nalu-Wind, through an overset technique from the TIOGA application. OpenFAST handles the structural dynamics of the turbine blades and towers, which informs the fluid-structure interaction of the wind turbines with the flow solvers. In the actuator line simulation, a mesh of 295M elements was used for a 5km x 5km domain, and it was simulated using 256 nodes (2048 GPU's) on the Oak Ridge Leadership Computing Facility Frontier supercomputer. For the blade-resolved simulation, 1.5B element mesh was used in the AMR-Wind background 5km x 5km domain, and 16M elements were used for each turbine in the Nalu-Wind domains, for a total of 1.7B elements. This was simulated using 384 nodes on Frontier, with each node using 56 cores for Nalu-Wind and 8 GPU cores. The data in this archive includes the turbine outputs from OpenFAST, 2D sampling planes from AMR-Wind, and full-field solution files from AMR-Wind and Nalu-Wind.

17 WIND ENERGY↗

Monitoring of ground water table depth and soil moisture at the Point Reyes field site

Ground water table (GWT) depth and soil moisture (SM) have been monitored at several locations at the Point Reyes field site (Californian coastal grassland) from 2021 to 2024. Monitoring is still on-going and data may be added to this archive at later time. The SM data have been acquired using Teros 12 Meter soil moisture sensors placed at 10, 30, 60 and 90 cm depth at 5 locations along a small hillslope. These sensors also collect soil temperature and bulk conductance. In addition, some collocated sensors provide pore pressure and Photochemical Reflectance Index (PRI). The GWT depth has been inferred from various type of Onset pressure transducers. The pressure measurements have been corrected for atmospheric pressure variations and sensor position relative to the ground surface to infer GWT depth, as well as with RTK GPS data to infer GWT elevation. The GWT data have been acquired at 5 distinct locations from 2020 to 2024 with the sensors placed at about 4 m depth. In addition, GWT data has been acquired for the 2023-2024 period with sensors located in 1 m deep shallow wells installed near each deeper well. This data is intended to evaluate possibly different dynamic in shallow (perched) and deep aquifer. The datasets are all provided in csv format. Please note that the interpretation of the GWT data needs to be done with consideration of environmental and well characteristics at the site and uncertainty in various variables. For more information on GWT and SM data, please contact the author.

54 ENVIRONMENTAL SCIENCES↗

Hyporheic-zone Processes and Stream Oxygen Dynamics: Insights from a Multiscale Reactive Transport Model: Modeling Archive

This archive contains the data and Python scripts required to reproduce the analyses and figures in the study: Gomez-Velez, J. D., Rathore, S. S., Cohen, M. J., & Painter, S. L. (2025). Hyporheic-zone Processes and Stream Oxygen Dynamics: Insights from a Multiscale Reactive Transport Model. Submitted to Water Resources Research. The analysis utilizes the subgrid model Advection Dispersion Equation with Lagrangian Subgrids (ADELS) implemented in the Advanced Terrestrial Simulator (ATS; https://amanzi.github.io/ats/stable/). In this case, the ATS and Amanzi versions are (1) ATS version 1.5.1_f5ba18f8 and (2) Amanzi version 1.6-dev_53444cca4. The repository includes a Jupyter Notebook and the necessary data (Pandas DataFrames stored as pickle files) to generate the figures for the manuscript. Additionally, it contains Python scripts to create ATS input files, run the ATS simulations, and post-process the results. Finally, it provides routines for parameter estimation using the Single-Station Metabolism (SSM) model with the Differential Evolution Adaptive Metropolis (DREAM) Markov Chain Monte Carlo (MCMC) algorithm with ZS enhancements (DREAM-ZS).

54 ENVIRONMENTAL SCIENCES↗

A customizable data management framework for high-repetition-rate high-energy-density science

The high-energy-density (HED) physics community is moving toward a new paradigm of high-repetition-rate (HRR) operation. To fully leverage the scientific power of HRR HED facilities, all of the components of each subsystem (laser, targetry, and performance diagnostics) must be connected and synchronized in a reliable and robust manner while the data acquired are tagged and archived in real time. To this end, GA has begun developing a generalized NoSQL-database framework, the MongoDB repository for information and archiving. An organizational strategy has been developed that shifts HED data organization from a shot-based to a diagnostic-based approach in order to increase archival and retrieval efficiency that lends itself to optimization applications. This work is a first step in pushing HRR HED science toward data management solutions that emphasize machine actionability and aim to stimulate community engagement to define data standards in HED science.

Instruments & Instrumentation↗

Machine Learning (ML) Classifier to Assist Metadata Creation

The Atmospheric Radiation Measurement (ARM) Data Center is responsible for the timely collection, archival, and curation of science data products. These products are freely available through an online data repository. Metadata creation is paramount for scientific users to find and access over seven petabytes of atmospheric science data. The hierarchical metadata structure allows users to search for information at both broad and narrow levels. This project aims to leverage 30 years’ worth of manually created metadata to enable machine predictions of broad-term classifications from narrow-term descriptions. These classification predictions would assist metadata coordinators with their term selections. This paper discusses the cleaning and preprocessing of the training data, the pipeline developed to determine the best model for this task, and the creation of an API metadata classifier for ARM measurement metadata. Our results show that the Linear Support Vector Classification (LinearSVC) algorithm, along with the Term Frequency – Inverse Document Frequency (TF-IDF) vectorizer, is well-suited for our multi-class classification task. Lengthier input training data led to better results, and artificial balancing was unnecessary for this particular use case. This predictive classifier enhances efficiency in metadata creation, as well as supports greater consistency and accuracy in metadata tagging.

Collier, Hannah [ORNL] (ORCID:0000000341284292)↗

Comparison of Radiosonde Datasets: SondeHub and Integrated Global Radiosonde Archive

SondeHub aggregates radiosonde telemetry data uploaded from community-run radiosonde receiver stations. This radiosonde telemetry dataset is open-source, available to anyone through Amazon S3. There are also other public radiosonde datasets such as National Centers for Environmental Information (NCEI)’s Integrated Global Radiosonde Archive (IGRA). While there are many similarities between the two datasets, there are many differences as well due to the nature of the two datasets: one is community-run, while the other is managed by a government agency. This report presents the result of analyzing and comparing the two datasets.

54 ENVIRONMENTAL SCIENCES↗

Impact of recent ENDF nuclear data update, high initial enrichment and high burnup fuel on critical experiments applicability determination via the integral index c k for burnup credit validation

In 2012, NUREG/CR-7109 reported on the validation of burnup credit calculations involving major and minor actinides and major fission products which was investigated for pressurized and boiling water reactor (PWR and BWR) fuel enrichments up to 5 wt% 235 U and assembly-average burnups up to 60 GWd/MTU. Recently, there has been interest in increasing the maximum enrichment used in PWR fuel as high as 8 wt% 235 U and correspondingly increasing the maximum assembly-average burnups to approximately 75 GWd/MTU. These proposed increases in enrichment and burnup necessitate reinvestigation of the validation basis for k eff calculations for this expanded application space. Additionally, the 2012 study was performed by using the Evaluated Nuclear Data File (ENDF)/B-VII.0 nuclear data with the SCALE 6 covariance library, and the effects of using the newly released ENDF/B-VII.1 and ENDF/B-VIII.0 nuclear data and covariance libraries should be evaluated. In this work, published in NUREG/CR-7309 in 2025, the validation assessment was performed consistently with NUREG/CR-7109: modeling irradiated fuel assemblies in the Generic Burnup Credit (GBC)-32 cask defined in NUREG/CR-6747. The TSUNAMI-3D sequence was used to generate sensitivity data for the application model, and the data were compared with sensitivity data from select benchmark models. The integral parameter c k is the metric of similarity used in this study and is consistent with NUREG/CR-7109, where a c k value in excess of 0.8 indicates sufficient similarity for use in validation. A new set of benchmark experiments with sensitivity data has been assembled for this effort. The number of experiments with available sensitivity data is now 2,104, compared to 474 in NUREG/CR-7109. This increase was facilitated by the efforts of the Nuclear Energy Agency to generate sensitivity data for a majority of the experiments in the International Criticality Safety Benchmark Evaluation Project (ICSBEP) Handbook to supplement the data available in the Oak Ridge National Laboratory (ORNL) Verified, Archived Library of Inputs and Data (VALID). The complete set of benchmarks considered here includes experiments for low-enriched uranium (LEU), intermediate enriched uranium (IEU), and a mixture of uranium and plutonium (MIX) from the ICSBEP Handbook and VALID, as well as ORNL models of the Haut Taux de Combustion (HTC) experiments and other potentially relevant models not included in VALID. The updated similarity study shows that none of the extended burnup and higher enrichment combinations considered show a significant decrease in the number of potentially applicable experiments, meaning sufficient critical experiments exist for the validation of BUC criticality safety calculations, with initial enrichments up to 8 wt% 235 U and burnups up to 80 GWd/MTU. Additionally, both the ENDF/B-VII.1 and ENDF/B-VIII.0 nuclear data libraries can be used for validation since the number of critical experiments applicable for validation increases for most cases with the most recent nuclear data compared to the previous one. As in previous BUC validation studies, the French HTC experiments are the most similar in a majority of the application cases studied, especially from representative discharge burnups ranging from 40 to 80 GWd/MTU. In conclusion, these results match the conclusions presented in NUREG/CR-7109 regarding validation of the primary actinides in BUC analyses.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Enhancing Discoverability and Management of Atmospheric Data at Scale: Solutions from the ARM Data Center

The Atmospheric Radiation Measurement (ARM) is a multi-laboratory and multi-institutional U.S. Department of Energy (DOE) Office of Science National User Facility. The ARM Data Center (ADC), located at Oak Ridge National Laboratory, collects, archives, and shares vast atmospheric data crucial for climate research. The ADC manages over 7 PB of data from 460 instruments worldwide, processing it into more than 11,000 diverse data products using the Network Common Data Form (NetCDF) for machine-independent accessibility. The primary challenge addressed in this paper is the efficient management and distribution of vast and diverse datasets essential for the climate research community, enhancing accessibility through advanced tools like Data Discovery. The ADC has developed advanced infrastructure and software architecture to handle the continuous influx of heterogeneous data to enhance data discoverability, resulting in increased scientific collaboration. In 2023, users from over 34 countries downloaded and utilized ARM data, resulting in 1,455 publications. The ADC’s efforts have significantly improved the discoverability and usability of atmospheric data, fostering extensive scientific research and collaboration. This paper details the solutions implemented by the ADC team for efficient data discovery and distribution, and it demonstrates ARM’s capability of staging processed data for scientific analysis.

Shah, Chirag [ORNL] (ORCID:0000000203145737)↗

Soil moisture and temperature from 2019 to 2024 along northeast- and southwest-facing hillslopes at the Lower Montane site in the East River Watershed, Colorado

Soil moisture, temperature, and electrical conductivity have been monitored at multiple depths (between 10 and 50 cm) at 4 locations along a northeast-facing slope and 3 locations on the opposite southwest-facing slope at the Lower Montane site in the East River Watershed, Colorado, from Oct 2019 to Oct 2024. The purpose of this data is to inform hydro-biogeochemical analyses for the Watershed Function Scientific Focus Area (SFA). Two locations on the northeast-facing slope were reinstalled in 2020 due to damage from wildlife, and thus data for these sites are provided in two distinct files. Overall, the data are reported in 9 CSV files containing the measurements, and the locations are provided in the Sensor_Location.csv file. There is a total of 10 *.csv data files and 3 *.csv metadata files. Older datasets associated with the northeast-facing slope are provided in another archive (see reference). These data products are part of the Watershed Function Scientific Focus Area collection effort to further scientific understanding of biogeochemical dynamics from genome to watershed scales. Feel free to contact the author with any questions or collaboration interests.

54 ENVIRONMENTAL SCIENCES↗

Electricity Baseline 2022 Background Data and Log File

The ElectricityLCI v2 Python package (https://github.com/USEPA/ElectricityLCI/tree/v2.0) was used to generate the 2022 electricity baseline: a regionalized life cycle inventory model of U.S. electricity generation, consumption, and distribution using standardized facility and generation data. ElectricityLCI implements a local data store for downloading and accessing public data on an individual's computer. The data store follows the folder definition provided by USEPA's esupy Python package (https://github.com/USEPA/esupy), which utilized the appdirs Python dependency (https://pypi.org/project/appdirs/). This submission includes the background data used to generate the 2022 electricity baseline inventory. Each zip archive stores the source files as found in their data stores. Sub-folders in each of the data stores are archived separately. For example, stewi.zip contains the JSON files, while stewi.facility.zip is the 'facility' sub-folder of stewi data store that stores the parquet files. To reproduce the data store, extract each zip file and drag-and-drop sub-folders in to their appropriate root folders to recreate the data stores, then copy the root folders to your data store folder (as returned by running the following on the command line: `python -c "import appdirs; print(appdirs.user_data_dir())"`). The main five data stores include: 'electricitylci', 'facilitymatcher', 'fedelemflowlist', 'stewi', and 'stewicombo'. The log file generated by the 2022 model run is also included, which contains the statements at the DEBUG level and above.

Electricity; LCA; data inventory↗

Electricity Baseline 2021 Background Data and Log File

The ElectricityLCI v2 Python package (https://github.com/USEPA/ElectricityLCI/tree/v2.0) was used to generate the 2021 electricity baseline: a regionalized life cycle inventory model of U.S. electricity generation, consumption, and distribution using standardized facility and generation data. ElectricityLCI implements a local data store for downloading and accessing public data on an individual's computer. The data store follows the folder definition provided by USEPA's esupy Python package (https://github.com/USEPA/esupy), which utilizes the appdirs Python dependency (https://pypi.org/project/appdirs/). An overview of the ElectricityLCI data stores may be found on the README (https://github.com/USEPA/ElectricityLCI/blob/v2.0/README.md#data-store). This submission includes the background data used to generate the 2021 electricity baseline inventory. Each zip archive stores the source files as found in their data stores. Sub-folders in each of the data stores are archived separately. For example, stewi.zip contains the JSON files, while stewi.facility.zip is the 'facility' sub-folder of stewi data store that stores the parquet files. To reproduce the data store, extract each zip file and drag-and-drop sub-folders in to their appropriate root folders to recreate the data stores, then copy the root folders to your data store folder (as returned by running the following on the command line: python -c "import appdirs; print(appdirs.user_data_dir())"). The main five data stores include: 'electricitylci', 'facilitymatcher', 'fedelemflowlist', 'stewi', and 'stewicombo'. The log file generated by the 2021 model run is also included, which contains the statements at the DEBUG level and above.

Electricity; LCA; LCI; Life Cycle; data inventory↗

Electricity Baseline 2020 Background Data and Log File

The ElectricityLCI v2 Python package (https://github.com/USEPA/ElectricityLCI/tree/v2.0) was used to generate the 2020 electricity baseline: a regionalized life cycle inventory model of U.S. electricity generation, consumption, and distribution using standardized facility and generation data. ElectricityLCI implements a local data store for downloading and accessing public data on an individual's computer. The data store follows the folder definition provided by USEPA's esupy Python package (https://github.com/USEPA/esupy), which utilizes the appdirs Python dependency (https://pypi.org/project/appdirs/). An overview of the ElectricityLCI data stores may be found on the README (https://github.com/USEPA/ElectricityLCI/blob/v2.0/README.md#data-store). This submission includes the background data used to generate the 2020 electricity baseline inventory. Each zip archive stores the source files as found in their data stores. Sub-folders in each of the data stores are archived separately. For example, stewi.zip contains the JSON files, while stewi.facility.zip is the 'facility' sub-folder of stewi data store that stores the parquet files. To reproduce the data store, extract each zip file and drag-and-drop sub-folders in to their appropriate root folders to recreate the data stores, then copy the root folders to your data store folder (as returned by running the following on the command line: python -c "import appdirs; print(appdirs.user_data_dir())"). The main five data stores include: 'electricitylci', 'facilitymatcher', 'fedelemflowlist', 'stewi', and 'stewicombo'. The log file generated by the 2020 model run is also included, which contains the statements at the DEBUG level and above.

Electricity; LCA; LCI; Life Cycle; data inventory↗

STM/S Grid LDOS Data and Analysis Code for Deciphering Majorana Zero Modes in Topological Superconductor

This dataset provides raw millikelvin scanning tunneling microscopy/spectroscopy (STM/S) grid spectroscopy data and Python analysis scripts supporting the manuscript “Deciphering Majorana Zero Modes in Topological Superconductor FeTe0.55Se0.45 with Machine-Learning-Assisted Spectral Deconvolution.” The dataset includes a raw grid spectroscopy file acquired on FeTe0.55Se0.45 at 40 mK under magnetic field, together with Python/Jupytext analysis scripts used for STM/S data processing, visualization, spectral deconvolution, Lorentzian peak fitting, feature extraction, machine-learning-assisted clustering, and figure generation. These files support the analysis of vortex-core local density of states and the identification of zero-bias-peak-related spectral components from complex in-gap states. The dataset is intended to provide a citable archival record of the data and analysis code associated with the published manuscript and to support transparency and reproducibility of the reported STM/S and machine-learning workflow.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Advances in Metallic Fuel Database Development and Data Qualification

The Fuels Irradiation and Physics Database (FIPD [1]) is a comprehensive repository of data and documents related to Uranium-Zirconium based metallic fuel test pins. This database stores operational conditions of these pins, calculated using a suite of Argonne National Laboratory analysis codes developed during the Integral Fast Reactor (IFR) program. Key calculated data include axial distributions of power, temperature, fluence, burnup, and isotopic densities. Additionally, the FIPD holds post-irradiation examination (PIE) data such as fission gas release, gas chemistry measurements, and axial distributions derived from profilometry, gamma scanning, and neutron radiography. Complementing these data is an extensive archive of documents related to various pins and experiments. These include raw PIE records, design details, safety analyses, and operational reports. More detail about FIPD can be found in ref. [2]. The database development is an ongoing effort covering metallic fuel experiments from the Experimental Breeder Reactor II (EBR-II) and the Fast Flux Test Facility (FFTF). The recent improvements to the database and the data QA status are summarized in this paper.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗