Search NASA⌕ Search

SEARCH · Search NASA

Results for “data retrieval”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Soil Temperature and Moisture within the Kougarok Fire Complex, Kougarok Road Mile Marker 86, Seward Peninsula, Alaska, 2019-2023

Daily averages of soil temperature and moisture measured once every hour at different heights located at Intensive Monitoring Stations within the Kougarok Fire Complex, Kougarok Road Mile Marker 86 site. Data were retrieved annually from 2019-2023. Package contains 21 *.CSV data files plus a file level metadata *.CSV, data dictionary *.CSV, data file inventory *.CSV, and sensor location site map *.JPG. Data files have header rows, NaN fields indicate invalid or missing data, and negative vertical offsets are above ground.The Kougarok tundra fire complex (KFC) is located north of Nome and the Kigluaik Mountains, near Quartz Creek and the Kougarok River. The site is accessed by foot from the end of the Nome-Taylor Highway (mile marker 86; also called the Kougarok or Beam Road). The KFC burned in six major fires in the decades since 1950 (Alaska Interagency Coordination Center, unpublished data). Lightning ignited five of these fires (1971, 1997, 2015, and 2019) and one was human caused (2002). The mosaic of overlapping fire scars allows for the study of repeat fires in the tundra which, until recently, was not a common phenomenon outside the boreal forest in Alaska. Our reference unburned tundra fire site is south of the KFC located at mile marker 80 of the Nome-Taylor Highway.The two most recent fires are the Mingvk Lake (2015; 21,698 acres burned from 7/27/2015 to 9/28/2015) and Garfield Creek (2019; 422 acres burned from 7/31/2019 to 8/20/19). The Mingvk Lake fire scar includes areas that burned 1-4x (1971, 1997, 2002), while the entirety of the Garfield Creek fire scar has burned 2x previously (1971, 2002).Previous research at the KFC focused on permafrost (Liljedahl et al. 2007; Narita et al. 2015; Iwahana et al. 2016; Tsuyuzaki, Iwahana, and Saito 2017) and vegetation (Narita et al. 2015; Hollingsworth et al. 2021) response to fire. The central Seward Peninsula is characterized by continuous permafrost with a thickness of 15 to 30 m and a mean active layer thickness of 56 cm (Hinzman et al. 2003). Sloping hills with mixed shrub–tussock tundra and tussock tundra vegetation in the uplands are characteristic of the region. Three micrometeorological towers near the Kougarok field site recorded a mean annual temperature of −2.4°C, mean January temperature of −23.1°C, mean July temperature of +11°C, and mean summer rainfall (June–August) of 94 mm from 2000 to 2006 (Liljedahl et al. 2007).The Next-Generation Ecosystem Experiments: Arctic (NGEE Arctic), was a research effort to reduce uncertainty in Earth System Models by developing a predictive understanding of carbon-rich Arctic ecosystems and feedbacks to climate. NGEE Arctic was supported by the Department of Energy's Office of Biological and Environmental Research.The NGEE Arctic project had two field research sites: 1) located within the Arctic polygonal tundra coastal region on the Barrow Environmental Observatory (BEO) and the North Slope near Utqiagvik (Barrow), Alaska and 2) multiple areas on the discontinuous permafrost region of the Seward Peninsula north of Nome, Alaska.Through observations, experiments, and synthesis with existing datasets, NGEE Arctic provided an enhanced knowledge base for multi-scale modeling and contributed to improved process representation at global pan-Arctic scales within the Department of Energy's Earth system Model (the Energy Exascale Earth System Model, or E3SM), and specifically within the E3SM Land Model component (ELM).

54 ENVIRONMENTAL SCIENCES↗

Evaluation of cloud height, optical thickness, and phase retrievals from the CHROMA algorithm applied to Sentinel-3 OLCI data

We previously developed the Cloud Height Retrieval from O 2 Molecular Absorption (CHROMA) algorithm for the Ocean Color Instrument (OCI) on the new NASA Plankton, Aerosol, Cloud, ocean Ecosystem (PACE) mission. Here, we apply CHROMA to observations from the Ocean Land Colour Instrument (OLCI) to guide expectations for PACE, as it will take some time to obtain large-scale validation data for OCI. We use cloud top height (CTH), phase, and (for liquid clouds) cloud optical thickness (COT) data from the ground-based Atmospheric Radiation Measurement (ARM) network to evaluate the OLCI retrievals. We found that OLCI and Moderate Resolution Imaging Spectroradiometer (MODIS) CTH compare similarly well to the ARM reference. OLCI has a tendency to underestimate CTH as CTH increases, and algorithm assumptions about cloud geometric thickness may contribute to this. ARM COT from multifilter shadowband radiometers (MFRSR) and Sun photometers are well-correlated with one another, albeit with a roughly 30 % offset on average; OLCI and MODIS COT agree more closely with the MFRSR data. OLCI retrieval uncertainty estimates show skill at telling low-uncertainty cases from high-uncertainty ones, although CTH uncertainties are underestimated. Additionally, we compare the OLCI data to satellite retrievals based on thermal infrared measurements from MODIS and Sea and Land Surface Temperature Radiometer (SLSTR) data. Differences are broadly consistent with physical expectations based on the A-band vs. thermal techniques, although one key challenge in such aggregated comparisons is different cloud masking sensitivities and algorithm failure rates meaning additional sampling differences are introduced. We conclude by discussing the transition to and possible enhancements for PACE OCI.

Sayer, Andrew M. [Univ. of Maryland Baltimore Coun↗

MCNPy

SAND2026-20425O MCNPy runs and analyzes simulations from MCNP, a software that models radiation transport of neutrons and gamma rays. MCNPy uses Python to start MCNP, retrieve event data files, and convert them into graph structures for detailed analysis. It offers visualization tools, including 2D views of particle histories, making complex simulation data easier to interpret for researchers and engineers. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy's National Nuclear Security Administration under contract DE-NA0003525.

Nowack, Aaron [Sandia National Lab. (SNL-CA), Live↗

Automating Log Synthesis and Visualization with Python and Splunk

The goal of this project is to automate log analysis by utilizing Splunk, Bash, and Python together. Simplifying the monitoring and analysis of network traffic was the main goal. In order to accomplish this, a Bash script was created to use 'tcpdump' to automate network sniffing. It also included a 24-hour file rotation mechanism to effectively manage the pcap files that were generated. After that, a Python script was written to read these pcap files and retrieve pertinent data about network traffic. After processing the collected data, Splunk is used to summarize the important metrics and visualize said information with relevant graphs.

99 GENERAL AND MISCELLANEOUS↗

Direct structural retrieval from gas-phase ultrafast diffraction data using a genetic algorithm

Ultrafast scattering techniques such as ultrafast electron diffraction and ultrafast x-ray diffraction have been utilized to elucidate the structural dynamics, reaction intermediates, and final products in molecular reactions following photoexcitation. The time-dependent structures are typically not directly retrieved from the experimental data, but they rely on comparison with calculations. The genetic algorithm (GA), a global optimization strategy, can be used to retrieve the molecular structures directly from diffraction patterns without any theoretical input. However, the robustness of the GA with respect to real experimental conditions such as a limited momentum transfer range, noise, and artifacts has not been studied in detail. In this work, we characterize the performance of the GA with simulated data that mimic realistic experimental conditions. We have developed and implemented a variant of the GA specific to diffraction measurements which performs better in the presence of imperfect data compared to the standard implementation of the GA. We demonstrate this method with both synthetic data and experimental ultrafast electron diffraction data on the UV-induced photodissociation of trifluoroiodomethane (C⁢F 3⁡ I) molecules.

74 ATOMIC AND MOLECULAR PHYSICS↗

Waveform retrieval for ultrafast applications based on convolutional neural networks

Electric field waveforms of light carry rich information about dynamical events on a broad range of timescales. The insight that can be reached from their analysis, however, depends on the accuracy of retrieval from noisy data. In this article, we present a novel approach for waveform retrieval based on supervised deep learning. We demonstrate the performance of our model by comparison with conventional denoising approaches, including wavelet transform and Wiener filtering. The model leverages the enhanced precision obtained from the nonlinearity of deep learning. The results open a path toward an improved understanding of physical and chemical phenomena in field-resolved spectroscopy.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

FY 2025 Multidimensional Data Correlation Platform: Unified Software Architecture for Advanced Materials and Manufacturing Technologies Data Management and Processing

The Advanced Materials and Manufacturing Technologies (AMMT) program continues to advance a data-driven approach to demonstrate the utility of additive manufacturing for fabricating components for nuclear applications. A key scientific goal is to leverage data to better understand manufacturing outcomes and thereby improve the performance, reliability, and lifespan of nuclear components. Ultimately, this effort supports the development of standards for certification and qualification of additively manufactured components, enabling broader industry adoption. In support of this objective, the AMMT program is building and deploying a data management platform to record, index, analyze, and make available the manufacturing data generated across the AMMT program. In FY 2023, the team conceptualized the architecture of the platform and, in FY 2024, deployed the first functional version at the Oak Ridge National Laboratory (ORNL) Manufacturing Demonstration Facility (MDF). In FY 2025, the platform was officially opened to all AMMT members. To enable this expansion, core modifications and enhancements were developed, including improvements to the user interface and workflows for data entry and retrieval. Most notably, robust security and access control mechanisms were implemented to protect data and manage information sharing. This effort featured a logging system, protected views, and controlled access mechanisms. This report documents these enhancements and the transition of the platform into program-wide use.

36 MATERIALS SCIENCE↗

Tansey_MICRE_D98_microphysics_V5

This data file contains retrievals of cloud-effective-radius and cloud-droplet-number-concentration for low clouds based on the Dong et al. 1998 retrieval technique for the MICRE campaign. The retrieval is based on observed downwelling broadband SW fluxes and microwave-radiometer-retrieved-liquid-water-path averaged on a 5 minute time scale (mean when low cloud is present). Details on the algorithm and analysis of results are given in Tansey et al. 2024 (submitted DOI: 10.22541/essoar.173482310.03809503/v1). The data is limited to the period 20160410 to 20160612 AND 20170103 to 20170214 during which time good quality microwave-radiometer measurements are available.

54 ENVIRONMENTAL SCIENCES↗

Tansey_MICRE_D98_microphysics_V5

This data file contains retrievals of cloud-effective radius and cloud-droplet-number concentration for low clouds based on the Dong et al. 1998 retrieval technique for the MICRE campaign. The retrieval is based on observed downwelling broadband shortwave (SW) fluxes and microwave-radiometer-retrieved-liquid-water-path averaged on a 5-minute time scale (mean when low cloud is present). Details on the algorithm and analysis of results are given in Tansey et al. 2024 (submitted DOI: 10.22541/essoar.173482310.03809503/v1). The data are limited to the period of 20160410 to 20160612 and 20170103 to 20170214 during which time good quality microwave radiometer measurements are available.

54 ENVIRONMENTAL SCIENCES↗

Search for a Cloud Phase Feedback in the Arctic Climate System

This project was motivated by a hypothesis involving the transition in lower troposphere temperatures across the freezing point of water. Specifically, at temperatures just below to about ten degrees below freezing, Arctic clouds should be in a mixed-phase, with strong influences from secondary ice production (e.g., the Hallett-Mossop process). At temperatures just above freezing, clouds should become entirely deglaciated. We hypothesized that, with multiple years of ARM data, a statistically significant change could be detected in cloud radiative properties and surface radiative fluxes that could be directly attributed to this phase change. Furthermore, in a gradually warming climate, these phase change-related responses in surface radiation would represent a cloud phase feedback as part of Arctic amplification. We designed this research project to coincide with a new ARM Arctic cloud radar product developed by Ed Luke and collaborators at BNL (Luke et al., 2021: PNAS, doi:10.1073/pnas.2021387118) that explicitly contains SIP cloud properties retrieved from radar data. This research project has successfully concluded after analysis of a much larger North Slope of Alaska (NSA) data sample than originally anticipated, and the results are quite different than originally hypothesized.

58 GEOSCIENCES↗

Variability of Eastern North Atlantic Summertime Marine Boundary Layer Clouds and Aerosols Across Different Synoptic Regimes Identified With Multiple Conditions

Abstract This study estimates the meteorological covariations of aerosol and marine boundary layer (MBL) cloud properties in the eastern North Atlantic (ENA) region, characterized by diverse synoptic conditions. Using a deep‐learning‐based clustering model with mid‐level and surface daily meteorological data, we identify seven distinct synoptic regimes during the summer from 2016 to 2021. Our analysis, incorporating reanalysis data and satellite retrievals, shows that surface aerosols and MBL clouds exhibit clear regime‐dependent characteristics, whereas lower tropospheric aerosols do not. This discrepancy likely arises from synoptic regimes determined by daily large‐scale conditions, which may overlook air mass histories that predominantly dictate lower tropospheric aerosol conditions. Focusing on three regimes dominated by northerly winds, we analyze the Atmospheric Radiation Measurement Program (ARM) ENA observations on Graciosa Island in the Azores. In the subtropical anticyclone regime, fewer cumulus clouds and more single‐layer stratocumulus clouds with light drizzle are observed, along with the highest cloud droplet number concentration (Nd), surface cloud condensation nuclei (CCN) and surface aerosol levels. The post‐trough regime features more broken or multi‐layer stratocumulus clouds with slightly higher surface rain rate, and lower Nd and surface CCN levels. The weak trough regime is characterized by the deepest MBL clouds, primarily cumulus and broken stratocumulus clouds, with the strongest surface rain rate and the lowest Nd, surface CCN and surface aerosol levels, indicating strong wet scavenging. These findings highlight the importance of considering the covariation of cloud and aerosol properties driven by large‐scale regimes when assessing aerosol indirect effects using observations.

54 ENVIRONMENTAL SCIENCES↗

Airborne imaging spectroscopy surveys of Arctic and boreal Alaska and northwestern Canada 2017–2023

Since 2015, NASA’s Arctic Boreal Vulnerability Experiment (ABoVE) has investigated how climate change impacts the vulnerability and/or resilience of the permafrost-affected ecosystems of Alaska and northwestern Canada. ABoVE conducted extensive surveys with the Next Generation Airborne Visible/Infrared Imaging Spectrometer (AVIRIS-NG) during 2017, 2018, 2019, and 2022 and with AVIRIS-3 in 2023 to characterize tundra, taiga, peatlands, and wetlands in unprecedented detail. The ABoVE AVIRIS dataset comprises ~1700 individual flight lines covering ~120,000 km 2 with nominal 5 m × 5 m spatial resolution. Data include individual transects to capture important gradients like the tundra-taiga ecotone and maps of up to 10,000 km 2 for key study areas like the Mackenzie Delta. The ABoVE AVIRIS surveys enable diverse ecosystem science, provide crucial benchmark data for validating retrievals from the PACE, PRISMA, and EnMAP satellite sensors and help prepare for the SBG and CHIME missions. This paper guides interested researchers to fully explore the ABoVE AVIRIS spectral imagery and complements our guide to the ABoVE airborne synthetic aperture radar surveys.

Miller, Charles E. [California Institute of Techno↗

extapi-acsys

Provides public APIs to the Fermilab control system. This service exposes several GraphQL endpoints for various, logical APIs that clients may use to retrieve control system data and, in some cases, make changes to the control system. This service is currently running on acsys-proxy.fnal.gov on port 8000 with the development instance on port 8001. The middle layer of the control system uses gRPCs for communications. The GraphQL resolvers of this service use various gRPC services to obtain the information that is returned. This uses the async-graphql and warp crates to provide GraphQL over http support. The resolvers use the tonic crate for gRPC client support.

Neswold, Rich [Fermi National Accelerator Laborato↗

Building a FAIR data ecosystem for incorporating single-cell transcriptomics data into agricultural genome to phenome research

Introduction The agriculture genomics community has numerous data submission standards available, but the standards for describing and storing single-cell (SC, e.g., scRNA- seq) data are comparatively underdeveloped. Methods To bridge this gap, we leveraged recent advancements in human genomics infrastructure, such as the integration of the Human Cell Atlas Data Portal with Terra, a secure, scalable, open-source platform for biomedical researchers to access data, run analysis tools, and collaborate. In parallel, the Single Cell Expression Atlas at EMBL-EBI offers a comprehensive data ingestion portal for high-throughput sequencing datasets, including plants, protists, and animals (including humans). Developing data tools connecting these resources would offer significant advantages to the agricultural genomics community. The FAANG data portal at EMBL-EBI emphasizes delivering rich metadata and highly accurate and reliable annotation of farmed animals but is not computationally linked to either of these resources. Results Herein, we describe a pilot-scale project that determines whether the current FAANG metadata standards for livestock can be used to ingest scRNA-seq datasets into Terra in a manner consistent with HCA Data Portal standards. Importantly, rich scRNA-seq metadata can now be brokered through the FAANG data portal using a semi-automated process, thereby avoiding the need for substantial expert curation. We have further extended the functionality of this tool so that validated and ingested SC files within the HCA Data Portal are transferred to Terra for further analysis. In addition, we verified data ingestion into Terra, hosted on Azure, and demonstrated the use of a workflow to analyze the first ingested porcine scRNA-seq dataset. Additionally, we have also developed prototype tools to visualize the output of scRNA-seq analyses on genome browsers to compare gene expression patterns across tissues and cell populations. This JBrowse tool now features distinct tracks, showcasing PBMC scRNA-seq alongside two bulk RNA-seq experiments. Discussion We intend to further build upon these existing tools to construct a scientist-friendly data resource and analytical ecosystem based on Findable, Accessible, Interoperable, and Reusable (FAIR) SC principles to facilitate SC-level genomic analysis through data ingestion, storage, retrieval, re-use, visualization, and comparative annotation across agricultural species.

Genetics & Heredity↗

Modeling performance of data collection systems for high-energy physics

Exponential increases in scientific experimental data are outpacing silicon technology progress, necessitating heterogeneous computing systems—particularly those utilizing machine learning (ML)—to meet future scientific computing demands. The growing importance and complexity of heterogeneous computing systems require systematic modeling to understand and predict the effective roles for ML. We present a model that addresses this need by framing the key aspects of data collection pipelines and constraints and combining them with the important vectors of technology that shape alternatives, computing metrics that allow complex alternatives to be compared. For instance, a data collection pipeline may be characterized by parameters such as sensor sampling rates and the overall relevancy of retrieved samples. Alternatives to this pipeline are enabled by development vectors including ML, parallelization, advancing CMOS, and neuromorphic computing. By calculating metrics for each alternative such as overall F1 score, power, hardware cost, and energy expended per relevant sample, our model allows alternative data collection systems to be rigorously compared. We apply this model to the Compact Muon Solenoid experiment and its planned high luminosity-large hadron collider upgrade, evaluating novel technologies for the data acquisition system (DAQ), including ML-based filtering and parallelized software. The results demonstrate that improvements to early DAQ stages significantly reduce resources required later, with a power reduction of 60% and increased relevant data retrieval per unit power (from 0.065 to 0.31 samples/kJ). However, we predict that further advances will be required in order to meet overall power and cost constraints for the DAQ.

Olin-Ammentorp, Wilkie (ORCID:0000000224729862)↗

ENDFtk: A robust tool for reading and writing ENDF-formatted nuclear data

ENDFtk is a recently developed C++ and Python interface to interact with ENDF-6 formatted nuclear data files. It provides a robust and complete interface, allowing the reading and writing of all formats currently part of the ENDF-6 formats manual, as well as some non-ENDF formats used by the NJOY processing code. It provides an interface that mimics the names in the ENDF-6 formats manual as well as an equivalent interface using human-readable attribute names. It is robust and powerful enogh for nuclear data experts to develop complex applications, while also simple enough to be used non-experts to retrieve and manipulate evaluated nuclear data. ENDFtk offers the ability to easily interrogate and manipulate data either in large-scale code projects or in simple Python scripts. Here, in this paper, a brief overview of the interface is given, as well as more substantial examples demonstrating plotting simple data, interacting with more complex data, and writing new data to files. ENDFtk is open source and available for download via GitHub (https://github.com/njoy/ENDFtk).

97 MATHEMATICS AND COMPUTING↗

A Comprehensive Northern Hemisphere Particle Microphysics Data Set From the Precipitation Imaging Package

Microphysical observations of precipitating particles are critical data sources for numerical weather prediction models and remote sensing retrieval algorithms. However, obtaining coherent data sets of particle microphysics is challenging as they are often unindexed, distributed across disparate institutions, and have not undergone a uniform quality control process. This work introduces a unified, comprehensive Northern Hemisphere particle microphysical data set from the National Aeronautics and Space Administration precipitation imaging package (PIP), accessible in a standardized data format and stored in a centralized, public repository. Data is collected from 10 measurement sites spanning 34° latitude (37°N–71°N) over 10 years (2014–2023), which comprise a set of 1,070,000 precipitating minutes. The provided data set includes measurements of a suite of microphysical attributes for both rain and snow, including distributions of particle size, vertical velocity, and effective density, along with higher-order products including an approximation of volume-weighted equivalent particle densities, liquid equivalent snowfall, and rainfall rate estimates. The data underwent a rigorous standardization and quality assurance process to filter out erroneous observations to produce a self-describing, scalable, and achievable data set. Case study analyses demonstrate the capabilities of the data set in identifying physical processes like precipitation phase-changes at high temporal resolution. Bulk precipitation characteristics from a multi-site intercomparison also highlight distinct microphysical properties unique to each location. This curated PIP data set is a robust database of high-quality particle microphysical observations for constraining future precipitation retrieval algorithms, and offers new insights toward better understanding regional and seasonal differences in bulk precipitation characteristics.

54 ENVIRONMENTAL SCIENCES↗

rcsb-api : Python Toolkit for Streamlining Access to RCSB Protein Data Bank APIs

The Protein Data Bank (PDB) was founded in 1971 as the first open-access digital data resource in biology to serve as the single global archive for three-dimensional (3D) macromolecular structure data. Current PDB holdings exceed 230,000 experimentally determined structures of proteins, nucleic acids, viruses, and macromolecular machines. The RCSB Protein Data Bank RCSB.org research-focused web portal facilitates search, analyses, and visualization of every PDB structure along with more than one million Computed Structure Models from AlphaFold DB and the ModelArchive. It is powered by a set of publicly available Application Programming Interfaces (APIs) that both support RCSB.org users and provide programmatic access to PDB data. Given the breadth and levels of granularity encompassed in this rich data collection, efficiently accessing the information programmatically may be challenging for new users. RCSB PDB has developed a Python software package, rcsb-api , that facilitates easy and efficient use of RCSB PDB APIs within a Python environment. This software tool is designed to streamline access to the extensive corpus of data housed within the PDB, enabling researchers to search, retrieve, and analyze 3D biostructure data seamlessly. Its use will accelerate research in structural biology, molecular biology and biochemistry, drug discovery, and bioinformatics by providing more efficient tools for data integration and analysis. The new toolkit is available on GitHub (github.com/rcsb/py-rcsb-api) and published to the public Python package repository (PyPI) to foster wider usage and support basic and applied research in fundamental biology, biomedicine, and the energy sciences.

FAIR principles↗