Search NASA⌕ Search

SEARCH · Search NASA

Results for “data sets”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 505 records · Page 28

SPRUCE Vegetation Phenology in Experimental Plots from PhenoCam Imagery, 2015-2024

This data set consists of PhenoCam data from the SPRUCE experiment from the beginning of whole ecosystem warming (Hanson et al. 2017) in August 2015 through March 31 of 2025 (2015-08-24 to 2025-03-31), with start- and end-of-season phenological transition dates derived through the end of autumn 2024. Digital cameras, or phenocams, installed in each SPRUCE enclosure track seasonal variation in vegetation “greenness”, a proxy for vegetation phenology and associated physiological activity. Three separate regions of interest (ROIs) were defined for each camera field of view, corresponding to different vegetation types and demarcating (1) Picea trees (vegetation type EN, for evergreen needleleaf); (2) Larix trees (vegetation type DN, for deciduous needleleaf); and (3) the mixed shrub layer (vegetation type SH). This data set consists of three sets of data files: (1) 3-day summary product files: One file for each camera and each ROI (i.e. vegetation type), characterizing vegetation color at a 3-day time step. • Contains 36 files in *.csv format inside a compressed (*.zip) file. (2) Transition date file: Estimates “greenness rising” (spring) and “greenness falling” (autumn) transition dates derived from the smoothed daily green chromatic coordinate (GCC) values, for each camera and each ROI (i.e., vegetation type). • Contains one file in *.csv format. (3) Snow flag files: Indicate days with snow on trees or snow on ground for each experimental enclosure. • Contains two files in *.csv format, one for snow on trees and one for snow on ground. This data set consists of two sets of companion files: (1) Accompanying HTML files show the 90th quantiles of the mean GCC plotted together with transition dates for each vegetation type and plot. • Contains three files in HTML format, one for each vegetation type. • One additional file in HTML format with the transition dates plotted for each vegetation type, by year. (2) R files for processing PhenoCam files and flags. • Contains five files in R file(*.R) format and the components of the phenocamr package (Version 1.1.4) used for calculating transition dates for 2015-2024. These are contained in a compressed (*.zip) file. User Note: All imagery is posted in near-real time to the PhenoCam Project web page (https://phenocam.nau.edu), where it is publicly available. Scroll to “spruce” in the Gallery or link directly to the 29 SPRUCE cameras at https://tinyurl.com/sprucecams. This data set is based on the complete camera record from SPRUCE and supersedes all previously released PhenoCam datasets (see Related Data Sets). The estimated transition dates for previously released datasets may differ slightly (in most cases, by ±3 days or less), because following standard PhenoCam processing protocols (Richardson et al. 2018, Scientific Data), smoothing and interpolation, outlier removal, and transition date estimation are always conducted using the full data record.

54 ENVIRONMENTAL SCIENCES↗

Data space volumes and classification optimization of SPOT and Landsat TM data

In order to compare the data space volume of SPOT XS and Landsat TM images, three data sets, i.e., a wetlands/agricultural data set, an agricultural data set, and a forest data set, are examined. The comparisons are made for the same geographic area. The data space volumes for Landsat TM (2, 3, and 4) are found to be 70 to 100 percent larger than the volumes for the SPOT XS images. It is suggested that the additional midinfrared bands contribute to the difference in data space volumes between Landsat TM and SPOT XS. The data space volumes for Landsat TM bands 3, 4, and 5 are more than an order of magnitude greater than the volumes for the three band SPOT XS data sets. The volumes of the six-band Landsat TM images are four orders of magnitude greater than the SPOT XS. The analysis of the data space volumes is used to optimize the computation time and minimize the storage requirements of a maximum-likelihood classification based on a look-up table.

Ahearn, Sean C.↗

Non-linear relationships between daily temperature extremes and US agricultural yields uncovered by global gridded meteorological datasets

Global agricultural commodity markets are highly integrated among major producers. Prices are driven by aggregate supply rather than what happens in individual countries in isolation. Furthermore, estimating the effects of weather-induced shocks on production, trade patterns and prices hence requires a globally representative weather data set. Recently, two data sets that provide daily or hourly records, GMFD and ERA5-Land, became available. Starting with the US, a data rich region, we formally test whether these global data sets are as good as more fine-scaled country-specific data in explaining yields and whether they estimate similar response functions. While GMFD and ERA5-Land have lower predictive skill for US corn and soybeans yields than the fine-scaled PRISM data, they still correctly uncover the underlying non-linear temperature relationship. All specifications using daily temperature extremes under any of the weather data sets outperform models that use a quadratic in average temperature. Correctly capturing the effect of daily extremes has a larger effect than the choice of weather data. In a second step, focusing on Sub Saharan Africa, a data sparse region, we confirm that GMFD and ERA5-Land have superior predictive power to CRU, a global weather data set previously employed for modeling climate effects in the region.

54 ENVIRONMENTAL SCIENCES↗

Pre-coding method and apparatus for multiple source or time-shifted single source data and corresponding inverse post-decoding method and apparatus

A pre-coding method and device for improving data compression performance by removing correlation between a first original data set and a second original data set, each having M members, respectively. The pre-coding method produces a compression-efficiency-enhancing double-difference data set. The method and device produce a double-difference data set, i.e., an adjacent-delta calculation performed on a cross-delta data set or a cross-delta calculation performed on two adjacent-delta data sets, from either one of (1) two adjacent spectral bands coming from two discrete sources, respectively, or (2) two time-shifted data sets coming from a single source. The resulting double-difference data set is then coded using either a distortionless data encoding scheme (entropy encoding) or a lossy data compression scheme. Also, a post-decoding method and device for recovering a second original data set having been represented by such a double-difference data set.

Yeh, Pen-Shu↗

Pre-coding method and apparatus for multiple source or time-shifted single source data and corresponding inverse post-decoding method and apparatus

A pre-coding method and device for improving data compression performance by removing correlation between a first original data set and a second original data set, each having M members, respectively. The pre-coding method produces a compression-efficiency-enhancing double-difference data set. The method and device produce a double-difference data set, i.e., an adjacent-delta calculation performed on a cross-delta data set or a cross-delta calculation performed on two adjacent-delta data sets, from either one of (1) two adjacent spectral bands coming from two discrete sources, respectively, or (2) two time-shifted data sets coming from a single source. The resulting double-difference data set is then coded using either a distortionless data encoding scheme (entropy encoding) or a lossy data compression scheme. Also, a post-decoding method and device for recovering a second original data set having been represented by such a double-difference data set.

Yeh, Pen-Shu↗

The relationship between aboveground biomass and radar backscatter as observed on airborne SAR imagery

The initial results of an experiment to examine the dependence of radar image intensity on total above-ground biomass in a southern US pine forest ecosystem are presented. Two sets of data are discussed. First, we examine two L-band (VV-polarization) data sets which were collected 5 years apart. These data sets clearly illustrate the change in backscatter resulting from the growth of a young pine stand. Second, we examine the dependence between radar backscatter and biomass as a function of radar frequency using data from the JPL Airborne Synthetic Aperture Radar (AIRSAR) and ERIM/NADC P-3 SAR systems. These results show that there is a positive correlation between above-ground biomass and radar backscatter and at C-, L-, and P-bands, but very little correlation at C-band. The biomass level for which this positive correlation holds decreases as radar frequency increases. This positive correlation is stronger at HH and HV polarizations that VV polarization at L- and P-bands, but strongest at VV polarization for C-band.

Kasischke, Eric S.↗

E-Standards For Mass Properties Engineering

A proposal is put forth to promote the concept of a Society of Allied Weight Engineers developed voluntary consensus standard for mass properties engineering. This standard would be an e-standard, and would encompass data, data manipulation, and reporting functionality. The standard would be implemented via an open-source SAWE distribution site with full SAWE member body access. Engineering societies and global standards initiatives are progressing toward modern engineering standards, which become functioning deliverable data sets. These data sets, if properly standardized, will integrate easily between supplier and customer enabling technically precise mass properties data exchange. The concepts of object-oriented programming support all of these requirements, and the use of a JavaTx based open-source development initiative is proposed. Results are reported for activity sponsored by the NASA Langley Research Center Innovation Institute to scope out requirements for developing a mass properties engineering e-standard. An initial software distribution is proposed. Upon completion, an open-source application programming interface will be available to SAWE members for the development of more specific programming requirements that are tailored to company and project requirements. A fully functioning application programming interface will permit code extension via company proprietary techniques, as well as through continued open-source initiatives.

Cerro, Jeffrey A.↗

Improved modeling of GPS selective availability

Selective Availability (SA) represents the dominant error source for stand-alone users of the Global Positioning System (GPS). Even for DGPS, SA mandates the update rate required for a desired level of accuracy in realtime applications. As was witnessed in the recent literature, the ability to model this error source is crucial to the proper evaluation of GPS-based systems. A variety of SA models were proposed to date; however, each has its own shortcomings. Most of these models were based on limited data sets or data which were corrupted by additional error sources. A comprehensive treatment of the problem is presented. The phenomenon of SA is discussed and a technique is presented whereby both clock and orbit components of SA are identifiable. Extensive SA data sets collected from Block 2 satellites are presented. System Identification theory then is used to derive a robust model of SA from the data. This theory also allows for the statistical analysis of SA. The stationarity of SA over time and across different satellites is analyzed and its impact on the modeling problem is discussed.

Braasch, Michael S.↗

The High Resolution Stereo Camera (HRSC) for the Lunar Scout 1 Mission

The High Resolution Stereo Camera (HRSC) is a planetary imaging system developed by the German Aerospace Research Establishment (DLR) with the involvement of the German Space Industry under the leadership of the German Space Agency (DARA) for the Russian Mars 94 and Mars 96 missions. The same instrument, virtually unmodified, is ideal for imaging the Moon. If flown on a Lunar Scout spacecraft, the HRSC will be operated so that it will produce data suitable for generation of a global lunar geodetic net, a global stereo image data set (both data sets produced at an orbit altitude of 200 kms approximately) and high resolution stereo imagery of areas of interest to the scientific community from an orbit altitude of 100 kms (resolution is a function of orbit altitude). All data will be digital.

Neukum, G.↗

Global Carbon Monoxide Products from Combined AIRS, TES and MLS Measurements on A-Train Satellites

This study tests a novel methodology to add value to satellite data sets. This methodology, data fusion, is similar to data assimilation, except that the background modelbased field is replaced by a satellite data set, in this case AIRS (Atmospheric Infrared Sounder) carbon monoxide (CO) measurements. The observational information comes from CO measurements with lower spatial coverage than AIRS, namely, from TES (Tropospheric Emission Spectrometer) and MLS (Microwave Limb Sounder). We show that combining these data sets with data fusion uses the higher spectral resolution of TES to extend AIRS CO observational sensitivity to the lower troposphere, a region especially important for air quality studies. We also show that combined CO measurements from AIRS and MLS provide enhanced information in the UTLS (upper troposphere/lower stratosphere) region compared to each product individually. The combined AIRS-TES and AIRS-MLS CO products are validated against DACOM (differential absorption mid-IR diode laser spectrometer) in situ CO measurements from the INTEX-B (Intercontinental Chemical Transport Experiment: MILAGRO and Pacific phases) field campaign and in situ data from HIPPO (HIAPER Pole-to-Pole Observations) flights. The data fusion results show improved sensitivities in the lower and upper troposphere (20-30% and above 20%, respectively) as compared with AIRS-only version 5 CO retrievals, and improved daily coverage compared with TES and MLS CO data.

combined AIRS↗

Development of global model for atmospheric backscatter at CO2 wavelengths

The improvement of an understanding of the variation of the aerosol backscattering at 10.6 micron within the free troposphere and the development model to describe this was undertaken. The analysis combines theoretical modeling with the results contained within three independent data sets. The data sets are obtained by the SAGE I/SAM II satellite experiments, the GAMETAG flight series and by direct backscatter measurements. The theoretical work includes use of a bimodal, two component aerosol model, and the study of the microphysical and associated optical changes occurring within an aerosol plume. A consistent picture is obtained, which describes the variation of the aerosol backscattering function in the free troposphere with altitude, latitude, and season. Most data are available and greatest consistency is found inside the Northern Hemisphere.

Kent, G. S.↗

NAIF Toolkit - Extended

The Navigation Ancillary Infor ma tion Facility (NAIF) at JPL, acting under the direction of NASA s Office of Space Science, has built a data system named SPICE (Spacecraft Planet Instrument Cmatrix Events) to assist scientists in planning and interpreting scientific observations (see figure). SPICE provides geometric and some other ancillary information needed to recover the full value of science instrument data, including correlation of individual instrument data sets with data from other instruments on the same or other spacecraft. This data system is used to produce space mission observation geometry data sets known as SPICE kernels. It is also used to read SPICE kernels and to compute derived quantities such as positions, orientations, lighting angles, etc. The SPICE toolkit consists of a subroutine/ function library, executable programs (both large applications and simple utilities that focus on kernel management), and simple examples of using SPICE toolkit subroutines. This software is very accurate, thoroughly tested, and portable to all computers. It is extremely stable and reusable on all missions. Since the previous version, three significant capabilities have been added: Interactive Data Language (IDL) interface, MATLAB interface, and a geometric event finder subsystem.

Acton, Charles H., Jr.↗

Ab Initio-Based Bond Order Potential for Arsenene Polymorphs Developed via Hierarchical Reinforcement Learning

Arsenene, a less-explored two-dimensional material, holds the potential for applications in wearable electronics, memory devices, and quantum systems. This study introduces a bond-order potential model with Tersoff formalism, the ML-Tersoff, which leverages multireward hierarchical reinforcement learning (RL), trained on an ab initio data set. This data set covers a spectrum of properties for arsenene polymorphs, enhancing our understanding of its mechanical and thermal behaviors without the complexities of traditional models requiring multiple parameter sets. Our RL strategy utilizes decision trees coupled with a hierarchical reward strategy to accelerate convergence in high-dimensional continuous search spaces. Unlike the Stillinger-Weber approach, which demands separate formalisms for buckled and puckered forms, the ML-Tersoff model concurrently captures multiple properties of the two polymorphs by effectively representing the local environment, thereby avoiding the need for different atomic types. Here, we apply the ML model to understand the mechanical and thermal properties of the arsenene polymorphs and nanostructures. We observe an inverse relationship between the critical strain and temperature in arsenene. Thermal conductivity calculations in nanosheets show good agreement with ab initio data, reflecting a decrease in thermal conductivity attributable to increased anharmonic effects at higher temperatures. We also apply the model to predict the thermal behavior of arsenene nanotubes.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Ligand-Based Compound Activity Prediction via Few-Shot Learning

Predicting the activities of new compounds against biophysical or phenotypic assays based on the known activities of one or a few existing compounds is a common goal in early stage drug discovery. This problem can be cast as a “few-shot learning” challenge, and prior studies have developed few-shot learning methods to classify compounds as active versus inactive. However, the ability to go beyond classification and rank compounds by expected affinity is more valuable. We describe Few-Shot Compound Activity Prediction (FS-CAP), a novel neural architecture trained on a large bioactivity data set to predict compound activities against an assay outside the training set, based on only the activities of a few known compounds against the same assay. Our model aggregates encodings generated from the known compounds and their activities to capture assay information and uses a separate encoder for the new compound whose activity is to be predicted. The new method provides encouraging results relative to traditional chemical-similarity-based techniques as well as other state-of-the-art few-shot learning methods in tests on a variety of ligand-based drug discovery settings and data sets.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Development of a global model for atmospheric backscatter at CO2 wavelengths

The variation of the aerosol backscattering at 10.6 micrometers within the free troposphere was investigated and a model to describe this variation was developed. The analysis combines theoretical modeling with the results contained within three independent data sets. The data sets used were obtained by the SAGE I/SAM II satellite experiments, the GAMETAG flight series, and by direct backscatter measurements. The theoretical work includes use of a bimodal, two component aerosol model, and the study of the microphysical and associated optical changes occurring within an aerosol plume. A consistent picture is obtained that describes the variation of the aerosol backscattering function in the free troposphere with altitude, latitude, and season.

Kent, G. S.↗

Analysis of positron lifetime spectra in polymers

A new procedure for analyzing multicomponent positron lifetime spectra in polymers was developed. It requires initial estimates of the lifetimes and the intensities of various components, which are readily obtainable by a standard spectrum stripping process. These initial estimates, after convolution with the timing system resolution function, are then used as the inputs for a nonlinear least squares analysis to compute the estimates that conform to a global error minimization criterion. The convolution integral uses the full experimental resolution function, in contrast to the previous studies where analytical approximations of it were utilized. These concepts were incorporated into a generalized Computer Program for Analyzing Positron Lifetime Spectra (PAPLS) in polymers. Its validity was tested using several artificially generated data sets. These data sets were also analyzed using the widely used POSITRONFIT program. In almost all cases, the PAPLS program gives closer fit to the input values. The new procedure was applied to the analysis of several lifetime spectra measured in metal ion containing Epon-828 samples. The results are described.

Singh, Jag J.↗

The topology of large-scale structure. III - Analysis of observations

A recently developed algorithm for quantitatively measuring the topology of large-scale structures in the universe was applied to a number of important observational data sets. The data sets included an Abell (1958) cluster sample out to Vmax = 22,600 km/sec, the Giovanelli and Haynes (1985) sample out to Vmax = 11,800 km/sec, the CfA sample out to Vmax = 5000 km/sec, the Thuan and Schneider (1988) dwarf sample out to Vmax = 3000 km/sec, and the Tully (1987) sample out to Vmax = 3000 km/sec. It was found that, when the topology is studied on smoothing scales significantly larger than the correlation length (i.e., smoothing length, lambda, not below 1200 km/sec), the topology is spongelike and is consistent with the standard model in which the structure seen today has grown from small fluctuations caused by random noise in the early universe. When the topology is studied on the scale of lambda of about 600 km/sec, a small shift is observed in the genus curve in the direction of a 'meatball' topology.

Gott, J. Richard, III↗

Accelerating Large Data Analysis By Exploiting Regularities

We present techniques for discovering and exploiting regularity in large curvilinear data sets. The data can be based on a single mesh or a mesh composed of multiple submeshes (also known as zones). Multi-zone data are typical to Computational Fluid Dynamics (CFD) simulations. Regularities include axis-aligned rectilinear and cylindrical meshes as well as cases where one zone is equivalent to a rigid-body transformation of another. Our algorithms can also discover rigid-body motion of meshes in time-series data. Next, we describe a data model where we can utilize the results from the discovery process in order to accelerate large data visualizations. Where possible, we replace general curvilinear zones with rectilinear or cylindrical zones. In rigid-body motion cases we replace a time-series of meshes with a transformed mesh object where a reference mesh is dynamically transformed based on a given time value in order to satisfy geometry requests, on demand. The data model enables us to make these substitutions and dynamic transformations transparently with respect to the visualization algorithms. We present results with large data sets where we combine our mesh replacement and transformation techniques with out-of-core paging in order to achieve significant speed-ups in analysis.

Moran, Patrick J.↗