Search NASA⌕ Search

SEARCH · Search NASA

Results for “Databases”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 253 records · Page 14

Classification of Cloud Particle Imagery and Thermodynamics (COCPIT): A New Databasing Tool for the Characterization of Cloud Particle Images Captured During DOE Field Campaigns

The Department of Energy for decades has explored the earth system and atmosphere through research and deployment of in-situ and remote sensing platforms during field campaigns. Among these datasets exists a vast supply of cloud particle images that provide visual insight into the complex microphysics in the clouds that span our globe. The millions of images collected over decades of deployments provides a unique opportunity to further our understanding of our atmosphere down to the crystal size. This work over the past 5 years has sought to organize these images into digestible datasets that can then be used by scientists to further our understanding of microphysics. A machine learning model was developed that categorizes over 1.5 million images across 11 weather events with over 90% accuracy according to particle type. The database was then extended to include dimensional characteristics of the particle as well as co-location of environmental properties, such as temperature and water content. Then, to initialize the connection between these data and our understanding of how crystals form and grow, weather research and forecasting simulations were run to generate the growth histories of the classified crystals. This research culminates with 2 databases per event: (1) a database of all classified crystals and their dimensional and environmental properties and (2) simulated growth histories of each crystal. Finally, a user interface was created to allow researchers to explore data statistics.

54 ENVIRONMENTAL SCIENCES↗

Efficient hemodynamic event detection utilizing relational databases and wavelet analysis

Development of a temporal query framework for time-oriented medical databases has hitherto been a challenging problem. We describe a novel method for the detection of hemodynamic events in multiparameter trends utilizing wavelet coefficients in a MySQL relational database. Storage of the wavelet coefficients allowed for a compact representation of the trends, and provided robust descriptors for the dynamics of the parameter time series. A data model was developed to allow for simplified queries along several dimensions and time scales. Of particular importance, the data model and wavelet framework allowed for queries to be processed with minimal table-join operations. A web-based search engine was developed to allow for user-defined queries. Typical queries required between 0.01 and 0.02 seconds, with at least two orders of magnitude improvement in speed over conventional queries. This powerful and innovative structure will facilitate research on large-scale time-oriented medical databases.

NASA Discipline Cardiopulmonary↗

A new version of the RDP (Ribosomal Database Project)

The Ribosomal Database Project (RDP-II), previously described by Maidak et al. [ Nucleic Acids Res. (1997), 25, 109-111], is now hosted by the Center for Microbial Ecology at Michigan State University. RDP-II is a curated database that offers ribosomal RNA (rRNA) nucleotide sequence data in aligned and unaligned forms, analysis services, and associated computer programs. During the past two years, data alignments have been updated and now include >9700 small subunit rRNA sequences. The recent development of an ObjectStore database will provide more rapid updating of data, better data accuracy and increased user access. RDP-II includes phylogenetically ordered alignments of rRNA sequences, derived phylogenetic trees, rRNA secondary structure diagrams, and various software programs for handling, analyzing and displaying alignments and trees. The data are available via anonymous ftp (ftp.cme.msu. edu) and WWW (http://www.cme.msu.edu/RDP). The WWW server provides ribosomal probe checking, approximate phylogenetic placement of user-submitted sequences, screening for possible chimeric rRNA sequences, automated alignment, and a suggested placement of an unknown sequence on an existing phylogenetic tree. Additional utilities also exist at RDP-II, including distance matrix, T-RFLP, and a Java-based viewer of the phylogenetic trees that can be used to create subtrees.

Non-NASA Center↗

Fatigue Crack Growth Database for Damage Tolerance Analysis

The objective of this project was to begin the process of developing a fatigue crack growth database (FCGD) of metallic materials for use in damage tolerance analysis of aircraft structure. For this initial effort, crack growth rate data in the NASGRO (Registered trademark) database, the United States Air Force Damage Tolerant Design Handbook, and other publicly available sources were examined and used to develop a database that characterizes crack growth behavior for specific applications (materials). The focus of this effort was on materials for general commercial aircraft applications, including large transport airplanes, small transport commuter airplanes, general aviation airplanes, and rotorcraft. The end products of this project are the FCGD software and this report. The specific goal of this effort was to present fatigue crack growth data in three usable formats: (1) NASGRO equation parameters, (2) Walker equation parameters, and (3) tabular data points. The development of this FCGD will begin the process of developing a consistent set of standard fatigue crack growth material properties. It is envisioned that the end product of the process will be a general repository for credible and well-documented fracture properties that may be used as a default standard in damage tolerance analyses.

Metallic materials↗

Image Database Exploration: Progress and Challenges

In this paper we discuss the general problem of automated image database exploration, the particular aspects of image databases which distinguish them from other databases, and how this impacts the application of off-the-shelf learning algorithms to problems of this nature. Current progress will be illustrated using two large-scale image exploration projects at JPL. The paper concludes with a discussion of current and future challenges.

image database exploration pattern recognition tec↗

Steps Toward Improved Integration, Search, and Analysis of Heterogeneous Data in the Astrobiology Habitable Environments Database

The Astrobiology Habitable Environments Database (AHED) is a new data system being developed as a long-term, open-access repository for astrobiology data. AHED is intended to store user-contributed results from NASA or externally-funded research in astrobiology, and to encourage sharing and synergy within the astrobiology community. However, the interdisciplinary nature of astrobiology presents some specific challenges to data management, integration, and analysis within AHED. In some disciplines (e.g., genomics), open databases thrive because the contributed products are fairly uniform and standardized (e.g., sequence data). In astrobiology, each investigation produces a unique set of data products; this makes it difficult to search across different datasets to find similar data, or to combine results from separate investigations. With AHED, we are taking steps to ensure there is adequate metadata - both at the dataset and record levels - to facilitate search, integration, and analysis. At the dataset level, we are developing a new metadata standard for describing astrobiology datasets, with detailed information about content, funding source, and scientific relevance, along with a set of topical keywords for characterizing datasets. At the record level, we are encouraging users to provide more structured content and finer-grained metadata. In many user-contributed science data repositories, few restrictions are placed on the uploaded data format, and minimal or no record-level metadata is required; thus users are unburdened when it comes to data preparation. The tradeoff is that deep integration and search across datasets is almost impossible without standardized structures and metadata. Although AHED users are free to upload minimally-described datasets, they will be encouraged to use database authoring tools (supplied by the underlying platform - Open Data Repository's Data Publisher) plus a set of customizable astrobiology-specific templates to help structure their data and provide standardized metadata. In reward for their extra effort, AHED will be able to deliver enhanced search, discovery, and analysis capabilities.

astrobiology↗

The Multiplatform Precipitation Feature (MPF) Database: Synthesizing Satellite and Ground-Based Precipitation and Lightning Datasets for Convective Studies

NASA’s Lightning Imaging Sensor (LIS) and the Global Precipitation Measurement (GPM) mission have contributed a wealth of data toward global lightning and precipitation studies, respectively. Combining lightning and precipitation datasets leverages their unique insights into deep convective processes that inform about characteristics of convection and its intensity. Recent efforts to synthesize the LIS and GPM datasets prepare the opportunity for unprecedented large-scale, value-added multiplatform analyses of convection. This data synthesis proof-of-concept study elaborates on the creation of a database of reflectivity-based multiplatform precipitation features (MPFs) that capture a combination of information extracted from spatiotemporally coincident lightning and precipitation data within individual storm features. The space-based GPM Dual-frequency Precipitation Radar (DPR) provides a record of precipitation data, while the GPM Validation Network (VN) additionally incorporates ground-based polarimetric Doppler radar data to provide microphysical and kinematic context to DPR data. The LIS instrument onboard the International Space Station has contributed lightning observations since 2017. MPFs encapsulating information from these datasets are created from isolated regions of filtered, smoothed DPR reflectivity data to which ellipses are fit. Each MPF includes feature location, size, and eccentricity information as well as summary reflectivity characteristics. They also include summaries of precipitation microphysics and derived three-dimensional wind available from ground-based radar data. LIS data provides standard lightning characteristics such as flash count and density to each MPF as well as other informative metrics such as flash area and radiance. Each MPF file includes information about the original data from which the MPF and its characteristics were determined, allowing end-user reconstruction of the ellipse and deeper “level I” analysis of captured data. This database of VN-LIS MPFs enables broad statistical analysis of the relationships between the microphysical, kinematic, and electrical properties of convection. Preliminary results from a demonstration of the database will be described as well as ongoing efforts and avenues for future work.

Lightning↗

On the New Optical Constants Database (OCdb) and its Importance for the Interpretation of Observational Data

The Optical Constants database(ocdb.smce.nasa.gov) came online in February 2023 and provides complex refractive indices of laboratory-generated organic refractory materials and ices relevant to (exo) planetary and astrophysical environments.The goal of the OCdb is to centralize published optical constants data to facilitate both their access by the scientific community and their use to analyze observational data returned by space missions and ground-based observatories. Computational tools are also under development to facilitate scientific use of the available OCdb optical constants data sets. Investigators generating laboratory optical constants are therefore encouraged to contribute their data to OCdb in order to increase the availability of their data and to enhance the scientific effectiveness of the database. Optical constants are critical input parameters in models (e.g.,radiative transfer, atmospheric, and reflectance spectral models)that are used to simulate the absorption, reflection, and scattering of light due to solid materials present in planetary and astrophysical environments (planets, their satellites, exoplanets, asteroids, comets, protoplanetary disks, etc.), and are key to the compositional interpretation of observations. We will first present the infrastructure of the OCdb and show how to use and contribute to it. We will introduce the two large NASA projects, namely, the Laboratory Astrophysics Directed Work Package and the NASA Center for Optical Constants, that have been instrumental in (i) developing the OCdb, (ii) generating planetary-and astrophysics-relevant ices and organic refractory materials from gas and ice irradiation in the laboratory, and (iii) determining their optical constants for inclusion in OCdb. We will also present two studies that are making use of these optical constants to interpret observations of Titan’s atmosphere and Pluto’s surface. These studies show the importance of measuring optical constants of laboratory-generated materials, and their impact on the models used to analyze and interpret astronomical observations. These studies also demonstrate the essential importance of such a database and the need for optical constants of a broad range of materials and wavelengths to enable the scientific community and to maximize the scientific return from space missions (e.g., Cassini, New Horizons, SOFIA, JWST).

The Optical Constants database↗

ECAR: Baseline Characterization Database Verification Report – PCEA Billet 01D3-35

The purpose of this engineering calculations and analysis report (ECAR) is to present data collected in the Baseline Graphite Characterization Program, which is directly tasked with supporting the Idaho National Laboratory’s (INL’s) research and development efforts on the Advanced Reactor Technologies (ART) Program. This program populates a comprehensive database that reflects the baseline properties of nuclear-grade graphite with regard to individual grade, billet, and position within individual billets. The physical- and mechanical-property information collected will be transferred to the Nuclear Data Management and Analysis System (NDMAS), and that database will help populate the handbook of property data available to member nations of the Generation-IV International Forum. Transfer of these data from the applicable technical lead to the dissemination databases available to other end users requires a full review of the test procedures and data-collection efforts through an analysis of the multiple summary spreadsheets and values being collected. This report represents the analysis for PCEA Billet 01D3-35 and facilitates release of associated data to the NDMAS custodians.

11 - NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

The ab initio non-crystalline structure database: empowering machine learning to decode diffusivity

Non-crystalline materials exhibit unique properties that make them suitable for various applications in science and technology, ranging from optical and electronic devices and solid-state batteries to protective coatings. However, data-driven exploration and design of non-crystalline materials is hampered by the absence of a comprehensive database covering a broad chemical space. In this work, we present the largest computed non-crystalline structure database to date, generated from systematic and accurate ab initio molecular dynamics (AIMD) calculations. We also show how the database can be used in simple machine-learning models to connect properties to composition and structure, here specifically targeting ionic conductivity. These models predict the Li-ion diffusivity with speed and accuracy, offering a cost-effective alternative to expensive density functional theory (DFT) calculations. Furthermore, the process of computational quenching non-crystalline structures provides a unique sampling of out-of-equilibrium structures, energies, and force landscape, and we anticipate that the corresponding trajectories will inform future work in universal machine learning potentials, impacting design beyond that of non-crystalline materials. In addition, combining diffusion trajectories from our dataset with models that predict liquidus viscosity and melting temperature could be utilized to develop models for predicting glass-forming ability.

36 MATERIALS SCIENCE↗

A global database of soil microbial phospholipid fatty acids and enzyme activities

Abstract Soil microbes drive ecosystem function and play a critical role in how ecosystems respond to global change. Research surrounding soil microbial communities has rapidly increased in recent decades, and substantial data relating to phospholipid fatty acids (PLFAs) and potential enzyme activity have been collected and analysed. However, studies have mostly been restricted to local and regional scales, and their accuracy and usefulness are limited by the extent of accessible data. Here we aim to improve data availability by collating a global database of soil PLFA and potential enzyme activity measurements from 12,258 georeferenced samples located across all continents, 5.1% of which have not previously been published. The database contains data relating to 113 PLFAs and 26 enzyme activities, and includes metadata such as sampling date, sample depth, and soil pH, total carbon, and total nitrogen. This database will help researchers in conducting both global- and local-scale studies to better understand soil microbial biomass and function.

Science & Technology - Other Topics↗

Geant4 Monte-Carlo (GEMC) A database-driven simulation program

GEMC[1] is an application that harnesses the power of databases to execute Geant4 Monte-Carlo simulations. The databases (MYSQL, CSQL, TEXT) define the geometry, materials, digitization algorithms, readout electronics and output formats. Implemented in C++, GEMC also boasts a user-friendly Python API that facilitates detector construction and database population. GEMC can handle real-life scenarios such as geometry variations and the run number-dependent calibration constants and digitization parameters. This abstract provides an overview of GEMC, accompanied by examples that showcase its versatility. We delve into the practical application of GEMC within the the CLAS12 experimental program at Jefferson Lab.

Ungaro, Maurizio↗

CABO-16S—a Combined Archaea, Bacteria, Organelle 16S rRNA database framework for amplicon analysis of prokaryotes and eukaryotes in environmental samples

Abstract Identification of both prokaryotic and eukaryotic microorganisms in environmental samples is currently challenged by the need for additional sequencing to obtain separate 16S and 18S ribosomal RNA (rRNA) amplicons or the constraints imposed by “universal” primers. Organellar 16S rRNA sequences are amplified and sequenced along with prokaryote 16S rRNA and provide an alternative method to identify eukaryotic microorganisms. CABO-16S combines bacterial and archaeal sequences from the SILVA database with 16S rRNA sequences of plastids and other organelles from the PR2 database to enable identification of all 16S rRNA sequences. Comparison of CABO-16S with SILVA 138.2 results in equivalent taxonomic classification of mock communities and increased classification of diverse environmental samples. In particular, identification of phototrophic eukaryotes in shallow seagrass environments, marine waters, and lake waters was increased. The CABO-16S framework allows users to add custom sequences for further classification of underrepresented clades and can be easily updated with future releases of reference databases. Addition of sequences obtained from Sanger sequencing of methane seep sediments and curated sequences of the polyphyletic SEEP-SRB1 clade resulted in differentiation of syntrophic and non-syntrophic SEEP-SRB1 in hydrothermal vent sediments. CABO-16S highlights the benefit of combining and amending existing training sets when studying microorganisms in diverse environments.

Eitel, Eryn M. (ORCID:0009000723919297)↗

Open database for GPD analyses

This article summarizes the main ideas behind creating an open database proposed for use in the exploration of generalized parton distributions (GPDs). This lightweight database is well suited for GPD phenomenology and is designed to store both experimental and lattice-QCD data. It can also aid in benchmarking GPD-related developments, such as GPD models. The database utilizes a new data format based on the YAML serialization language, enabling the storage of essential information for modern analyses, such as replica values. It includes interfaces for both Python and C++, allowing straightforward integration with analysis codes.

Burkert, V. D. [Thomas Jefferson National Accelera↗

Downloadable Dynamometer Database (D3): Public Test Data on Advanced-Technology Vehicles

Access to high-quality, independent vehicle test data is critical to advancing energy-efficient transportation research. The Downloadable Dynamometer Database (D3) is a public repository of dynamometer test data on advanced-technology vehicles, generated at the Advanced Mobility Technology Laboratory (AMTL) at Argonne National Laboratory and hosted by the Transportation and Power Systems Division. The database has been made available to support researchers, students, and professionals engaged in energy-efficient vehicle research, development, and education. A wide range of vehicle categories has been tested (i.e., alternative fuel vehicles, conventional gasoline and diesel vehicles, all-electric vehicles, hybrid electric vehicles, and plug-in hybrid electric vehicles), as well as various drive cycles and test conditions documented in the accompanying D3 user presentation. Stakeholders can select a vehicle type, identify a vehicle of interest, and download the associated test data for use in their own analyses. Data downloaded from D3 must be accompanied by the required attribution: "This data is from the Downloadable Dynamometer Database and was generated at the Advanced Mobility Technology Laboratory (AMTL) at Argonne National Laboratory." These data are critical to vehicle modeling, validation, technology assessment, and educational use.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

The Global Spectra-Trait Initiative: A database of paired leaf spectroscopy and functional traits associated with leaf photosynthetic capacity (v1.0.0)

The Global Spectra-Trait Initiative (GSTI) aims to generate generalizable spectra trait models using reflectance data to predict leaf traits associated with the photosynthesis capacity of leaves. It comprises a synthesized dataset of leaf trait data, input datasets and code. Leaf traits include the maximum carboxylation rate of rubisco (Vcmax), the maximum electron transport rate (Jmax), the dark respiration, as well as the prediction of leaf nitrogen, leaf mass per area (LMA), and leaf water content (LWC). The dataset comprises >7500 paired observations from around 400 species from a broad range of biomes. This dataset comprises a zip file of the GSTI GitHub repository (https://github.com/plantphys/gsti), the synthesized database (.csv) and database metadata files. This dataset was updated on 2025-12-12 with minor edits to mirror the accepted manuscript version and GitHub release (Version 1.0.0 (ESSD accepted version)). Edits included minor changes to the project documentation on GitHub and removal of 12 duplicate entries from the database.

54 ENVIRONMENTAL SCIENCES↗

Basin-Scale Structural Features Database

The Basin-Scale Structural Features database provides spatial datasets of faults, fractures, folds, and earthquakes compiled from public, authoritative sources (e.g., U.S. Geological Survey and State Geological Surveys) and aggregated into derivative forms to support subsurface assessments. Recognizing that characterizing basin-scale structural features requires interpreting data that are often ambiguous or lack key information, the source data were evaluated using a knowledge-data framework and geospatial fuzzy logic method (Justman et al., 2020) to represent both measured (observed) and predicted (inferred or potential) structural features as derivative datasets. This workflow employs conceptual models for known structural features and predicted structural features, incorporating geospatial data to estimate potential, even with limited data. The aim is to aid and support an understanding of basin-scale features and identify potential gaps in data and knowledge. As of 4/30/2025, the database includes resources for nine sedimentary basins: Appalachian, Denver, U.S. Gulf Coast, Illinois, Michigan, Permian, Sacramento, San Joquin and Williston. The database is organized by basin and then data category: 1) Faults, fractures, folds, 2) Earthquakes, 3) Topographic, 4) Structural contours and isopachs, 5) Geophysical, and 6) Structural feature density assessment maps.

basin scale↗