Search NASA⌕ Search

SEARCH · Search NASA

Results for “file format”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

DICOMs, Missiles, and Metadata: The U.S. Nuclear Weapons Program Leverages a Medical Standard

Digital Imagine and Communications in Medicine (DICOM) images, although most commonly used in medical settings, have been widely adopted by the United States Department of Energy (DOE) for capturing images of the internal components in a nuclear weapon. DICOMs, a lesser-known file format that combine nested metadata structures with complex image data-including multiple planes, frames, and high resolution-require the creation of access copies to support usability within the DOE. Used primarily for ensuring the safety, security, and reliability of the U.S. nuclear stockpile, the Los Alamos National Laboratory (LANL)’s DICOM images and corresponding image metadata must be accessible to scientists and researchers via our institutional centralized databases. This paper describes the author's creation of a Python script that converts DICOM images into accessible, archive-friendly TIFF files while preserving key image data and metadata.

96 KNOWLEDGE MANAGEMENT AND PRESERVATION↗

Total Radioisotope Yield Calculator

An initial number of fissions on a fissile/fissionable material, defined by ENDF library availability, is used to produce an initial inventory of fission products (FPs). The FPs are decayed using full, analytical solutions to the Bateman equations to calculate the FP inventory at a user-specified time. Output can contain information about activity, mass, dose rate, and gamma-ray intensities based on user-defined parameters, given as .txt or .xlsx file formats. Full execution and display of results takes 5-10 seconds, depending on the complexity of the input parameters.

Holschuh, ThomasV [Idaho National Laboratory (INL)↗

polars-dovmed (dovmed) v0.1.0

polars-dovmed is a python package for text search and extraction from NCBI's PubMed Central Open Access subset. It is powered by the polars dataframe library and leverages modern file formats (parquet) to efficiently scan public literature.

Roux, Simon [Lawrence Berkeley National Laboratory↗

ExaGO v2

ExaGO is a high-performance computing power systems modeling suite providing models for different power flow analyses. It supports forward AC power flow, multiperiod AC and DC optimal power flow analyses, contingency analysis, as well as stochastic optimal power flow analysis. ExaGO can use HiOp and Ipopt optimization engines. It supports Matpower and PSS/E input file formats. ExaGO v2 includes code from ExaGO 1.6.0.

Peles, Slaven [Oak Ridge National Laboratory (ORNL↗

Data for "Event Horizon Telescope Pattern Speeds in the Visibility Domain”

Visibility Amplitude Pattern Speeds from the v3 Sgr A* Library. Data for "Event Horizon Telescope Pattern Speeds in the Visibility Domain” (Conroy et al.). Data are provided in 2 file formats: a TXT table which is a standard format for the Astrophysical Journal (ApJ) where the paper is submitted and the original NPY format.

Conroy, Nicholas S. [University of Illinois] (ORCI↗

Utah FORGE: Well 16B(78)-32 Distributed Temperature Sensing Data from April and May 2024

This dataset includes Neubrex Energy Services fiber optic distributed temperature sensing (DTS) data from well 16B(78)-32 during stimulation and circulation, including interaction with well 16A(78)-32, during April and May 2024. The DTS data are stored in HDF5 file format and are accompanied by a PowerPoint report on the study. All times in this dataset are in UTC. Depths are in MD relative to Kelly Bushing Height, and temperatures are in degrees Fahrenheit. All DTS measurements were made using a Yokogawa 3000DTSX Distributed Temperature Sensing Interrogator Unit, with a spatial sampling interval of 3.28 feet and a temporal sampling rate of 129 seconds. The third-party Pressure-Temperature Gauge data should be used with caution after April 20, 2024, as its performance is not considered reliable beyond this date.

15 GEOTHERMAL ENERGY↗

Utah FORGE: Injection and Production Test results and Reports from August 2024

This dataset includes results and reports from the injection/production test for wells 16A(78)-32 and 16B(78)-32 in August, 2024 at Utah FORGE. Materials include injection, post stimulation production, flow rate, gamma, temperature and pressure survey results. For the respective wells, injection and production profile results and interpretations are provided in .xlsx and .las file formats. Additionally, for both wells, preliminary and final reports are included and provide logging procedures and visualizations of the results.

15 GEOTHERMAL ENERGY↗

Oscilloscope Data Push Program

Data acquisition (DAQ) is a complex and costly process. Creating DAQ systems for analyzing a system requires expensive electronics and a dedicated team of engineers for support, posing a challenge for users who readily need data. This project is a proof of concept to create a temporary or one-off DAQ system using equipment commonly available to every team. We aim to automate the data acquisition process from the Rohde \& Schwarz RTO 1044 oscilloscope, convert the acquired binary data into floating point values, and store the results in a CSV file format. By developing a Python program to handle these tasks, we seek to reduce the manual effort involved in data collection, significantly increasing efficiency.

Osei-Tutu, Jason↗

FY 2024 Multidimensional Data Correlation Platform Data Management Infrastructure Progress: Materials Laboratory

This report provides an inventory of the equipment available at the ORNL Manufacturing Demonstration Facility (MDF) for sample preparation and material characterization, including both destructive and non-destructive techniques that generate critical data to support the development of the Multi-Dimensional Data Correlation (MDDC) framework. The success of the MDDC framework depends heavily on the quality and completeness of the data it can access. Therefore, it is essential to establish a comprehensive inventory of the technologies available to the Advanced Materials and Manufacturing Technologies (AMMT) multi-laboratory team. This starts by gathering information about the types of data they produce, the data collection and transfer protocols used, file formats, and data storage requirements for experiments. This information is then carefully evaluated to create the operations and trackables elements of the Damara Tern platform, which is the foundation of the MDDC framework.

36 MATERIALS SCIENCE↗

Monthly Quality-filtered Aggregation of NOAA Climate Data Record (CDR) of AVHRR Leaf Area Index (LAI) and Fraction of Absorbed Photosynthetically Active Radiation (FAPAR), Version 5

This dataset contains gridded monthly Leaf Area Index (LAI) derived from the daily NOAA Climate Data Record (CDR) of AVHRR Leaf Area Index (LAI) and Fraction of Absorbed Photosynthetically Active Radiation (FAPAR), Version 5. This data record spans from 1981 to 2018 using data from eight NOAA polar orbiting satellites: NOAA-7, -9, -11, -14, -16, -17, -18 and -19. The data are projected on a 0.05 degree x 0.05 degree global grid, as in the original CDR. The original CDR is one of the Land Surface CDR Version 5 products produced by the NASA Goddard Space Flight Center (GSFC) and the University of Maryland (UMD), which is accompanied by algorithm documentation, data flow diagram and source code for the NOAA CDR Program. This dataset is in the netCDF-4 file format following ACDD and CF Conventions. This dataset has applied quality assurance information to only include "OK" data from the original CDR in the monthly aggregation.

Vermote, Eric [NASA Goddard Space Flight Center (G↗

Monthly Quality-filtered Aggregation of NOAA Climate Data Record (CDR) of AVHRR (Version 5) and VIIRS (Version 1) Leaf Area Index (LAI) and Fraction of Absorbed Photosynthetically Active Radiation (FAPAR)

This dataset contains gridded monthly Leaf Area Index (LAI) derived from the daily NOAA Climate Data Record (CDR) of AVHRR (Version 5) and VIIRS (Version 1) Leaf Area Index (LAI) and Fraction of Absorbed Photosynthetically Active Radiation (FAPAR). This data record spans from 1981 to 2024 using data from NOAA polar orbiting satellites: NOAA-7, -9, -11, -14, -16, -17, -18, -19 and S-NPP. The data are projected on a 0.05 degree x 0.05 degree global grid, as in the original CDR. The original CDR is one of the Land Surface CDR products produced by the NASA Goddard Space Flight Center (GSFC) and the University of Maryland (UMD), which is accompanied by algorithm documentation, data flow diagram and source code for the NOAA CDR Program. This dataset is in the netCDF-4 file format following ACDD and CF Conventions. This dataset has applied quality assurance information to only include "OK" data from the original CDR in the monthly aggregation.

Vermote, Eric [NASA Goddard Space Flight Center (G↗

Bias Corrected NOAA HRRR Wind Resource Data for Grid Integration Applications

To address the need for regularly updated wind resource data, NREL has processed the High-Resolution Rapid Refresh (HRRR) outputs for use in grid integration modeling. The HRRR is an hourly-updated operational forecast product produced by the National Oceanic and Atmospheric Administration (NOAA) (Dowell et al., 2022). Several barriers have prevented the HRRR's widespread proliferation in the wind energy industry: missing timesteps (prior to 2019), challenging file format for wind energy analysis, limited vertical height resolution, and negative bias versus legacy WIND Toolkit data (2007-2013). NREL has applied re-gridding, interpolation, and bias-correction to the native HRRR data to overcome these limitations. This results in the now-publicly-available bias corrected and interpolated HRRR (BC-HRRR) dataset for weather years 2015 to 2023. Bias correction is necessary for wind resource consistency across weather years to be used simultaneously in planning-focused grid integration studies alongside the original WIND Toolkit data. We show that quantile mapping with the WIND Toolkit as a historical baseline is an effective method for bias correcting the interpolated HRRR data: the BC-HRRR has reduced mean bias versus comparable gridded wind resource datasets (+0.12 m/s versus Vortex) and has very low mean bias versus ground measurement stations (+0.01 m/s) (Buster et al., 2024). BC-HRRR's consistency with the legacy WIND Toolkit allows NREL to extend grid integration analysis to 15+ weather years of wind data with low-overhead extensibility to future years as they are made available by NOAA. As with historical datasets like the WIND Toolkit, BC-HRRR is intended for use in grid integration modeling (e.g., capacity expansion, production cost, and resource adequacy modeling) both independently and alongside the legacy WIND Toolkit.

Array↗

BSEC VPRM 10m Hourly Biogenic Fluxes in Baltimore (2021)

Model outputs from the Vegetation Photosynthesis and Respiration Model (VPRM: version from Horne et al. in prep). Model remote sensing inputs come from Sential 2-derived EVI and LSWI. Model meteorological inputs for two-meter air temperature and shortwave incoming come from the BSEC WRF 2021 Control Run (Foust, W. 2023). Plant functional Types (PFTs) are spatially classified using the Chesapeake Bay Program 2018 land use land cover product. The final biogenic flux (µmol CO2 m^-2 s^-1) outputs of NEE, RESP, and GEE are a weighted average based on the portion of PFTs within the cell. Individual PFT outputs are saved inside PFT directories (e.g., Crops, Grass, etc.) inside the specific month directory. Model outputs are denoted as a negative flux into the land system (i.e., photosynthesis) and a positive flux as a net release into the overlying atmosphere. Respiration (RESP) fluxes are positive and combine heterotrophic (only soil) and autotrophic sources. Gross ecosystem exchange (GEE) is a negative flux driven by only photosynthetic activity from vegetation, and the Net ecosystem exchange (NEE) is the sum of the two (i.e., NEE=RESP+GEE). Data Characteristics Spatial Resolution: 10m Temporal Resolution: Hourly File Format: VPRM_ _BSEC. .tif (Hour is in UTC) For more information on the model results, please email Jason Horne (jph6488@psu.edu). References: Foust, W. (2023). BSEC WRF 2021 Control Run Output (v0.1.0) [Data set]. MSD-LIVE Data Repository. https://data.msdlive.org/records/m0e6m-vvq17

Baltimore↗

BSEC ecohydrological and water quality fluxes from RHESSys Simulations in USGS gauged watersheds

Baltimore Environmental Social Collaborative (BSEC) Water and Water Quality Simulations from RHESSys Model The repository contains RHESSys (Tague & Band, 2004; source code) simulated ecohydrological and nutrient (nitrogen only) fluxes at daily, basin-average (RHESSys_basin_output) and monthly, grid (RHESSys_patch_output) levels. We currently simulated the following 8 watersheds in Baltimore: Dead Run Baisman Run Scotts Level Branch Moores Run Powder Mill Run Maidens Choice Run Stony Run The watershed boundaries of all studied watersheds are stored in Watershed_Boundary folder. Variables and their units are listed in the metadata. Spatial projection, NAD83 / UTM zone 18N (EPSG:26918) is used for patch-level, netCDF-format files. For more information, please contact Ruoyu Zhang (rz3jr@virginia.edu).

Baltimore MD↗

Data and code from: Multivariate bayesian regression model for predicting disposed ash composition at U.S. coal fired power stations

This dataset contains the code and data files needed for implementation of a Multivariate Bayesian Regression model, described in Jin et al. (2025), for the historical prediction of the chemical composition of disposed coal ash at U.S. coal fired power plants as a function of annualized coal purchase data. The integrated coal supply data file (CoalSupplyDataset.csv) represents a compilation of monthly fuel purchase records for the period 1973-2022 at major U.S. power stations. These records were obtained from the U.S. Energy Information Administration. The CSV file also contains, for each coal purchase record, the coal region of the mine as defined by the U.S. Geological Survey. Data entry errors and data gaps in the EIA records were corrected as described in Jin et al. This CSV file represents the integrated coal supply data after corrections were made. The model structure and fitting parameters are encoded in pickle file format (Bayesian.pkl). The model was developed with the coal supply data and coal ash composition data, apportioned according to the Stratified Shuffle Split for training and testing subsets. The model was built using Python and the PyMC library. Reference Publication: Jin, Z.; Huang, J.; Hower, J.C.; Hsu-Kim, H.(2025). Predictive Assessment of the Chemical Composition of Coal Ash in Reserve at U.S. Disposal Sites. Environmental Science & Technology.

Coal ash composition↗

Oscilloscope Data Push Program

This paper details the development of a Python program designed to automate the data acquisition and conversion for an oscilloscope for the purposes of a one-off/temporary data acquisition system for users that readily need data, and do not have the option of obtaining a Data Acquisition (DAQ) solution. Creating DAQ systems for analyzing a system requires expensive electronics and a dedicated team of engineers for support. Traditionally, manual data collection and processing are time consuming and prone to error. By automating these processes, the cost, efficiency and accuracy of data handling are improved upon. This project involves the creation of a program that interacts with the oscilloscope. During this interaction, there are various functions being performed such as the acquisition of waveform data via floating points, generating plots with the acquired wave points, and storing of floating points in a CSV file format for future reference and plotting purposes. While the initial aim of the project included continuous logging to a cloud database, this was deferred due to time constraints. The results portrayed an almost-instant rate of data collection with a buffer time, showcasing the potential for further integration and real-time data processing.

Osei-Tutu, Jason↗

On the Abuse and Detection of Polyglot Files

A polyglot is a file that is valid in two or more formats. Polyglot files pose a problem for file-upload and generative AI web interfaces that rely on format identification to determine how to securely handle incoming files. In this work we found that existing file-format and embedded-file detection tools, even those developed specifically for polyglot files, fail to reliably detect polyglot files used in the wild. To address this issue, we studied the use of polyglot files by malicious actors in the wild, finding 30 polyglot samples and 15 attack chains that leveraged polyglot files. Using knowledge from our survey of polyglot usage in the wild---the first of its kind---we created a novel data set based on adversary techniques. We then trained a machine learning detection solution, PolyConv, using this data set. PolyConv achieves a precision-recall area-under-curve score of 0.999 with an F1 score of 99.20% for polyglot detection and 99.47% for file-format identification, significantly outperforming all other tools tested. We developed a content disarmament and reconstruction tool, ImSan, that successfully sanitized 100% of the tested image-based polyglots, which were the most common type found via the survey. Our work provides concrete tools and suggestions to enable defenders to better defend themselves against polyglot files, as well as directions for future work to create more robust file specifications and methods of disarmament.

Oesch, T [ORNL] (ORCID:0000000269091022)↗

Data-model files associated with the manuscript "Modeling the Effects of Wetland Restoration on Coastal Hydrology: A Case Study of Elkhorn Slough Watershed, California"

This package contains the data, simulation setups, notebooks and figures used in “Modeling the Effects of Wetland Restoration on Coastal Hydrology: A Case Study of Elkhorn Slough Watershed, California” (Xu et al., 2025). In this study, we selected Elkhorn Slough, a tidal estuary, in California, to investigate the impact of wetland restoration and sea level rise on coastal hydrology using the process-based coastal hydrologic model, Advanced Terrestrial Simulator (ATS), informed by site-specific data. We designed a novel modeling workflow for incorporating wetland restoration features into land cover and soil properties for the model parameterization. The validation results demonstrate a strong agreement between modeled and observed data. We studied the characteristics of coastal watershed hydrology, then focused on the surface water dynamics at two wetland sites within Elkhorn Slough, a reference site and a restored site. Our simulation results indicate that the restored site successfully maintains surface elevation, resulting in reduced surface inundation. We also examined the impact of wetland restoration under expected sea level rise over the next few decades. The low-lying Yampah Marsh, the reference site, is likely to be inundated due to future sea level rise when highest tides arrive; while a higher percentage of Hester Marsh, the restored site, would retain marsh vegetation in coming decades, regardless of tidal conditions. Our study provides important information for examining the outcome of restoration practices that include surface elevation in tidal wetlands under climate changes.Several files can be found from this data package.1. README.md: This file describes the title, journal, co-authors, abstract, repository structure and model version.2. Simulation_Setups.zip: The file contains the model configuration files (XML format) for ATS. 3. Notebooks.zip: The file contains the Jupyter notebooks for generating the pre- and post-restoration meshes and the meshes of future scenarios. 4. Figures.zip: The file contains the figures used in the manuscript.5. Data.zip: The file contains the data used to drive the model simulations, including watershed and wetlands boundaries, mesh files and references to additional datasets (e.g., meteorological forcing, tidal dataset, DEMs, land cover, soil properties). Also, it contains water level observations at the restored wetland.

54 ENVIRONMENTAL SCIENCES↗