Search NASASearch

SEARCH · Search NASA

Results for “database management systems”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

SoK: What does it Mean to Benchmark Database Forensics?

Relational Database Management Systems are the backbone of modern enterprises and public-sector services, and are thus frequent targets of security incidents, insider threats, and thorough regulatory audits. Consequently, databases have become key sources of digital evidence, requiring investigators to reconstruct past activity from audit logs, transaction logs, and backups. Although benchmarking frameworks such as those developed by the Transaction Processing Performance Council (TPC) are widely used to evaluate database performance, they do not capture forensic requirements such as evidentiary completeness, tamper-evidence, chain of custody, or regulatory compliance under GDPR and CCPA. This survey examines the emerging domain of forensic database benchmarking. We gathered prior research on database forensics, secure logging, and tamper-evident data structures; we analyze modern forensic-ready features in commercial and open-source systems (SQL Server Ledger, Oracle Blockchain Tables, PostgreSQL pgAudit, Db2 Audit, Aurora Database Activity Streams, Oracle Real Application Security and IBM Guardium) and assess why existing benchmarks are insufficient. We propose forensic workloads, metrics, and methodologies that incorporate adversarial stressors, deleted-record recovery, and backup analysis. We also identify open research problems and call for a community-driven forensic benchmark suite. The result is an idea for evaluating not only database performance but also forensic soundness, bridging the gap between system engineering, compliance, and digital investigations.

Lenard, Ben

Software and computing for Run 3 of the ATLAS experiment at the LHC

The ATLAS experiment has developed extensive software and distributed computing systems for Run 3 of the LHC. These systems are described in detail, including software infrastructure and workflows, distributed data and workload management, database infrastructure, and validation. The use of these systems to prepare the data for physics analysis and assess its quality are described, along with the software tools used for data analysis itself. An outlook for the development of these projects towards Run 4 is also provided.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS

Baseline Characterization Database Verification Report ? NBG-17 Billet V104

The purpose of this report is to present data collected in the Baseline Graphite Characterization Program, which is directly tasked with supporting the Idaho National Laboratory’s (INL’s) research and development efforts on the Advanced Reactor Technologies (ART) Program. This program populates a comprehensive database that reflects the baseline properties of nuclear-grade graphite regarding individual grade, billet, and position within individual billets. The physical- and mechanical-property information being collected will be transferred to the Nuclear Data Management and Analysis System (NDMAS), and that database will help populate the handbook of property data available to member nations of the Generation-IV International Forum. Transfer of these data from the applicable technical lead to the dissemination databases available to other end users requires a full review of the test procedures and data-collection efforts through an analysis of the multiple summary spreadsheets and values being collected. This report represents the analysis for NBG-17 Billet V104 and facilitates release of associated data to the NDMAS custodians. Millions of raw data points have been collected during testing and quantification analyses for these billets. The summary scalar property values and supplementary traceability data are collected into comprehensive spreadsheets. Data sets are composed of single billets of graphite for any given grade, organized by mechanical test-specimen type, and further subdivided into individual spreadsheet tabs according to the specific test or evaluation being performed. A direct analysis of properties was not conducted, and this report does not provide information on the validity or performance characteristics of the graphite itself. Rather, this report is intended as a verification of the completeness of actual data collected in accordance with PLN-3467, “Baseline Graphite Characterization Plan: Electromechanical Testing,” [1] and PLN-3348 “Graphite Mechanical Testing” [2] and their representation of the measurement and test results with sole regard to the graphite billets under evaluation.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS

ECAR: Baseline Characterization Database Verification Report – PCEA Billet 01D3-35

The purpose of this engineering calculations and analysis report (ECAR) is to present data collected in the Baseline Graphite Characterization Program, which is directly tasked with supporting the Idaho National Laboratory’s (INL’s) research and development efforts on the Advanced Reactor Technologies (ART) Program. This program populates a comprehensive database that reflects the baseline properties of nuclear-grade graphite with regard to individual grade, billet, and position within individual billets. The physical- and mechanical-property information collected will be transferred to the Nuclear Data Management and Analysis System (NDMAS), and that database will help populate the handbook of property data available to member nations of the Generation-IV International Forum. Transfer of these data from the applicable technical lead to the dissemination databases available to other end users requires a full review of the test procedures and data-collection efforts through an analysis of the multiple summary spreadsheets and values being collected. This report represents the analysis for PCEA Billet 01D3-35 and facilitates release of associated data to the NDMAS custodians.

11 - NUCLEAR FUEL CYCLE AND FUEL MATERIALS

Path Forward: Materials Data Modernization for ASME Codes and Standards in the Artificial Intelligence Era

Development of the ASME Materials Properties Database was initiated in the early 2010s to support the ASME Codes and Standards. As information technologies advance at an accelerated pace with the artificial intelligence era on the horizon, the ASME Materials Properties Database must be further modernized from a database to a knowledgebase to ride the wave of digital information revolution and effectively support the ASME Codes and Standards in the new era. This paper is intended to provide an overview of the ASME Materials Properties Database and discuss a roadmap for its future development to facilitate understanding of and participation from different sectors of the Codes and Standards community. Further, it first reviews the basic concepts of data, information, knowledge, database, and database system as well as the pros and cons in different types of data management and then discusses the path forward for a desired evolution of the database into a self-explanatory and machine-readable knowledgebase that is consistent with human cognitive processes for the Codes and Standards development and, furthermore, provides resources for data processing and analysis to reach an eventual goal of streamlining the Codes and Standards development from the initial inquiry, throughout data submission, analysis, …, to Codes and Standards rule establishment for final publication.

36 MATERIALS SCIENCE

Catalyzing deep decarbonization with federated battery diagnosis and prognosis for better data management in energy storage systems

Industrial data analytics methods play a central role in improving energy storage performance and efficiency, impacting the future of electrified transportation and renewable electricity generation. However, significant challenges hinder the large-scale deployment of batteries. Conventional methods rely on centralized collection and processing of fleet-level data, leading to database size issues and privacy concerns due to potential data breaches. To enable scalable deployment of battery management systems, this article proposes a federated battery diagnosis and prognosis model, which distributes the processing of battery standard current-voltage-time-usage data in a privacy-preserving manner. Instead of transferring the raw data, this approach communicates only the locally processed parameters, thus reducing communication load and preserving data confidentiality. The federated model offers a paradigm shift in battery health management through privacy-preserving distributed methods for battery data processing and lifetime prediction, ensuring the reliable and sustainable deployment of lithium-ion batteries in a rapidly evolving world.

asset health management

Untargeted, tandem mass spectrometry (LC/MS-MS) metaproteomes from soil samples in control and warming plots in Blodgett Forest, CA (2014-2021)

The pathways of carbon transport and loss through and from soils—soil organic matter (SOM) depolymerization to dissolved organic carbon and mineralization to carbon dioxide (CO2)—are fundamentally driven by microbial activity, which is strongly regulated by environmental conditions. As part of Lawrence Berkeley National Laboratory (LBNL) Terrestrial Ecosystem Science (TES) Belowground Biogeochemistry Science Focus Area (SFA), we have established a novel whole-soil long-term warming experiment at the University of California (UC) Blodgett Forest Research Station (Sierra Nevada) in 2014, where we study the role of biogeochemical, microbial and geochemical process interactions in SOM decomposition and stabilization. This package contains soil metaproteomics data in the context of site specific metagenomes from soil depth profiles in three paired control and warming plots from a temperate mixed forest in Northern California. Each paired plot had been subjected to experimental warming since June 2014 to simulate a predicted climate change scenario for northern California. These metaproteomes were collected in 2018 after 4.5 years of warming from five depth intervals (0-10 cm, 10-30 cm, 30-45 cm, 45-60 cm, 60-80 cm). For protein identification, the collected spectra were searched following a target-decoy search strategy against a database of metagenome predicted proteins (covering 96 samples from 2014 to 2021) representing the complete sequence diversity at the site. Data was searched with mass spectrometry database search tool (MS-GF+) using Pacific Northwest National Laboratory (PNNL)'s Data Management System (DMS) Processing pipeline. The metagenomes are published as part of another data package. Raw metaproteomic data and the data products from MS-GF+ are deposited in the Mass Spectrometry Interactive Virtual Environment (MassIVE) database under accession no. MSV000097826. Here we present a dataset that includes spectral counts for the detected proteins across samples (EMSL50964_BrodieAllMAGs_Globals_SC.txt), the sequences of the detected proteins, and sample metadata file that contains site information for the soil metaproteome samples.

Belowground Biogeochemistry Science Focus Area

An Integrated ML/AI Framework for Digitizing, Structuring and Searching DOE U-TRU-Fuels Data with Gap Analysis of Non-DOE Records

The U.S. Department of Energy (DOE) Advanced Fuels Campaign (AFC) is advancing transmutation fuel technologies to reduce long-lived radioactive waste by converting minor actinides into shorter-lived or stable elements through irradiation in sodium-cooled fast reactors. Key experiments such as AFC-1, AFC-2, FUels for the transmutation of Trans-URanium elements In phéniX (FUTURIX)-Fortes Teneurs en Actinides (FTA), and Experimental Breeder Reactor-II (EBR-II) X501 have provided fuel fabrication, irradiation, and performance data on various transuranic-bearing fuel forms. This report documents the creation of an artificial-intelligence assisted database, which has consolidated all DOE-owned data related to Transuranic (TRU)-bearing fuel experiments and stored across it across both the Idaho National Laboratory (INL) Nuclear Data Management and Analysis System and the INL high performance computing (HPC) infrastructure. A dedicated webpage, hosted on the INL HPC system, has been developed to support role-based access and data interaction. The database architecture allows researchers to navigate large, heterogeneous archives with far greater speed and accuracy than manual search and lays the foundation for future expansion into multimodal nuclear materials analysis environments. The database represents a major step towards a nationally integrated fuels database utilizing artificial intelligence tools.

11 - NUCLEAR FUEL CYCLE AND FUEL MATERIALS

Anthromes and forest carbon responses to global change

Human effects on ecosystems date back thousands of years, and anthropogenic biomes—anthromes—broadly incorporate the effects of human population density and land use on ecosystems. Forests are integral to the global carbon cycle, containing large biomass carbon stocks, yet their responses to land use and climate change are uncertain but critical to informing climate change mitigation strategies, ecosystem management, and Earth system modeling. Using an anthromes perspective and the site locations from the Global Forest Carbon (ForC) Database, we compare intensively used, cultured, and wildland forest lands in tropical and extratropical regions. We summarize recent past (1900-present) patterns of land use intensification, and we use a feedback analysis of Earth system models from the Coupled Model Intercomparison Project Phase 6 to estimate the sensitivity of forest carbon stocks to CO 2 and temperature change for different anthromes among regions. Modeled global forest carbon stock responses are positive for CO 2 increase but neutral to negative for temperature increase. Across anthromes (intensively used, cultured, and wildland forest areas), modeled forest carbon stock responses of temperate and boreal forests are less variable than those of tropical forests. Tropical wildland forest areas appear especially sensitive to CO 2 and temperature change, with the negative temperature response highlighting the potential vulnerability of the globally significant carbon stock in tropical forests. The net effect of anthropogenic activities—including land-use intensification and environmental change and their interactions with natural forest dynamics—will shape future forest carbon stock changes. These interactive effects will likely be strongest in tropical wildlands.

54 ENVIRONMENTAL SCIENCES

Modernizing GlideinWMS Factory Monitoring with Prometheus & Grafana

Large-scale scientific experiments like CMS and DUNE rely on the distributed workload management system GlideinWMS to efficiently utilize computing resources across heterogeneous computing environments. GlideinWMS currently records Factory statistics using Round Robin Databases (RRDBs), XML, and JSON files, and these statistics are displayed via custom monitoring Web pages, thereby limiting integration with modern observability platforms. This project investigates the use of Prometheus-based instrumentation to expose Factory metrics using OpenTelemetry principles. Factory statistics related to Glidein submission and job execution are exported as Prometheus metrics through the Prometheus Python Client Library and are served via an HTTP metrics endpoint. The collected metrics are inspected using the Prometheus web-based interface and are visualized through Grafana dashboards within the Landscape monitoring infrastructure at Fermilab. This project significantly streamlines the integration of modern monitoring technologies into GlideinWMS and establishes a framework for extending observability across additional system components.

Appiah, Gideon [Grambling State U.]

Capturing Historic Reliability Performance Through Graph Databases: A Model Based System Engineering Approach

With the goal of improving the performance and reliability of high dependable technological systems such as nuclear power plants, advanced monitoring and health management systems are employed to inform system engineers on observed degradation processes and anomalous behaviors of assets and components. This information is captured in the form of large amount of data which can be heterogenous in nature (e.g., numeric, textual). Such large data availability poses challenges when system engineers are required to parse and analyze them in order to track historic reliability performance of assets and components. This paper tackles directly this challenge by providing means to organize data in the form of a graph: a knowledge graph. The presented approach distinguish itself from current knowledge graph-based methods by the fact that model-based system engineering (MBSE) models are used to “put data into context”. In particular, MBSE models are used as skeleton of a knowledge graph; numeric and textual data elements, once processed, are associated to MBSE model elements. Thus, a knowledge graph captures both system architecture (though MBSE models) and health/performance data. Such feature opens the door to new data analytics methods designed to identify causal relations between observed phenomena.

97 - MATHEMATICS AND COMPUTING

Integrated Energy-Water Data for Cross-Sector Resilience

This white paper focuses on the “energy-for-water” domain, addressing the urgent need for integrated, empirical data to support regional management, benchmarking, and research on improving efficiency and developing technologies for water and wastewater management systems. The costs and energy required for the supply, treatment, and distribution of water and wastewater lack a standard data collection mechanism and centralized database or storage infrastructure, limiting data-driven decision-making across interdependent infrastructure systems.

42 ENGINEERING

Used Nuclear Fuel Management Using the Next Generation System Analysis Model

The U.S. Department of Energy (DOE) is leading the National effort to manage the back end of the nuclear fuel cycle, encompassing the safe transportation, storage/staging, and/or eventual disposal of used nuclear fuel (UNF) and high-level radioactive waste. The Next Generation System Analysis Model (NGSAM) is DOE’s discrete-event, agent-based simulation tool designed to model the full life cycle of UNF from reactor discharge to final disposal. NGSAM supports the DOE Office of Spent Fuel and High-Level Waste Disposition by enabling a detailed, scenario-based analysis of logistics, infrastructure, and shipping strategies. NGSAM replaces legacy models with a modern, flexible platform built on Repast Simphony and enhanced by the Process Analysis Tool. NGSAM simulates the movement and interaction of individual fuel assemblies with system components such as canisters, casks, railcars, and facilities. The model integrates with the Java Transportation Operations Model to plan and execute transportation scenarios, supporting both constrained and unconstrained resource allocation. Key features include customizable allocation and acceptance algorithms, detailed facility-level operations, and a Quick Edit tool for rapid scenario adjustments. NGSAM supports multimodal transportation modeling (e.g. rail, road, barge) and provides comprehensive cost, schedule, and infrastructure data. NGSAM utilizes data from sources such as DOE’s STANDARDS UNF database and DOE’s Stakeholder Tool for Assessing Radioactive Transportation, while also allowing user-defined inputs for scenario customization. NGSAM enables stakeholders to evaluate complex UNF management strategies, assess system performance under varying assumptions, and inform decision making for future infrastructure investments. Its modular architecture and integration with other Integrated Waste Management System tools make it a critical asset for planning the safe and efficient disposition of the Nation’s growing UNF inventory.

Craig, Brian [Argonne National Laboratory (ANL)]

Osprey Framework v0.2.2

The Alpha Berkeley Framework is a software architecture for building agentic AI systems that coordinate multi-step workflows in scientific and industrial environments. It is based on a plan-first orchestration model, where natural language requests are translated into execution plans with explicit dependencies and optional human approval. The framework includes capability classification, which selects relevant tools on a per-task basis to keep orchestration efficient as the number of available tools grows. It incorporates task extraction methods that compress conversational context and integrate external resources such as databases, APIs, and knowledge bases into structured, machine-readable tasks. Execution is supported by modular services with checkpointing, artifact management, and error handling, allowing workflows to be paused, inspected, and resumed. The system is designed for deployment in production environments, supporting both local and containerized execution as well as integration with HPC clusters. Interfaces include command-line tools, browser-based workflows, and containerized services. The framework has been demonstrated in tutorial examples and deployed at the Advanced Light Source, where it coordinates accelerator control and analysis workflows.

Hellert, Thorsten [Lawrence Berkeley National Labo

Cataloging Legacy Data from the Tritium Systems Test Assembly Program

The Tritium Systems Test Assembly (TSTA) at Los Alamos National Laboratory, operational from 1984 to 2001, was critical in advancing fusion fuel cycle technologies, including tritium storage, gas separation, and pumping. TSTA’s contributions, particularly in safe tritium operations, have influenced subsequent fusion projects. This paper discusses the ongoing effort to digitize and catalog TSTA’s historical data to create a searchable resource for the fusion research community. While the long-term objective is to develop a relational database for structured data management, the project remains in the early phase, with current efforts focused on scanning and indexing physical documents. Initial plans for database implementations are also presented, outlining key considerations for structure, query indexing, and standardization. As digitization progresses, future discussions will refine these implantation details to ensure an efficient and comprehensive system. This initiative aims to preserve critical legacy data, enhance the design of tritium system facilities, and support the next generation of fusion energy research.

42 ENGINEERING

M3SF-24LL010302062-NEA-TDB Management and International Collaborations in Sorption and Thermodynamic Modeling

This progress report (Level 3 Milestone Number M3SF-24LL010302062) summarizes research conducted at Lawrence Livermore National Laboratory (LLNL) within the Crystalline International Collaborations Work Package Number SF-24LL01030206. The activity is focused on our long-term commitment to engaging our partners in international nuclear waste repository research. This includes participation in the Nuclear Energy Agency Thermochemical Database (NEA-TDB) Project and development of methodologies for integrating US and international thermodynamic databases for use in SFWST Generic Disposal System Assessment (GDSA) efforts. A continuing focus for FY24 efforts is to support the US participation in the NEA-TDB effort. The focus of FY24 activities was the development of an agreement for a Phase 7 activity that will start in Q1 of 2025. Mavrik Zavarin is now the US representative on both the Management Board and the Executive group to the NEA-TDB. He is also the POC for the Cements State of the Art Report that is undergoing peer review in FY24. In FY24, we used our position on the NEA-TDB MB and EG to facilitate the integration of NEA-TDB thermochemical data with LLNL’s SUPCRTNE thermodynamic database that supports the SFWST GDSA activities. This effort is coordinated with the Argillite work package SUPCRTNE database development efforts (Wolery, 2024). The goal is to provide a downloadable database that will be hosted on LLNL’s thermodynamics website which incorporates NEA-TDB data into the LLNL database where appropriate. We also began engagement with the EURAD2 program that was initiated in FY24 by our European collaborators at the Karlsruhe Institute of Technology (KIT), Germany. The primary focus of the engagement is with WP20: DITUSC Thermodynamic database evaluation program. A kickoff meeting for this activity is planned for early FY25. Finally, we have been selected to co-host (with Clemson University) the International Conference on Chemistry and Migration Behaviour of Actinides and Fission Products in the Geosphere in 2025 (Migration2025). The meeting will be held September 21-26, 2025, in New Orleans, Louisiana, and will focus on international efforts to understand the risks of radionuclide releases into the environment. This central focus of this conference is on international efforts to develop safe disposal options for nuclear wastes. As such, we are developing a theme focused on US underground nuclear waste repository science.

58 GEOSCIENCES

An Open-source Llm Enhanced-tool Specialized In Helping Moose Related Problems And Tasks

MOOSEenger is an open-source, terminal-first chat application for the MOOSE ecosystem that couples specialized parsing of MOOSE documentation and “.i” input files with retrieval-augmented generation to deliver grounded answers about multiphysics modeling and workflows. It includes dedicated readers for MOOSE-style HTML and a pyhit-based parser that uses the MOOSE syntax tree to preserve block structure and attach retrieval metadata. A data-ingestion pipeline performs semantic chunking into atomic facts and stores them hierarchically in a local Chroma vector database that maintains parent–child relationships across documents; the system can ingest directories, individual files, and single-page web content, and it provides CRUD operations (insert, update, delete) to manage the corpus. At query time, relevant chunks are embedded, retrieved, and fused into the model context, with interactive features such as token streaming, persistent chat history, and dynamic RAG (retrieval triggered by user input or intermediate model output). Deployment is flexible: MOOSEenger runs with local Ollama models or remote Hugging Face/OpenAI backends—typically coordinating generation, lightweight tagging/summarization, and embeddings across three models—and it also supports a server mode and integration with the VS Code Continue interface.

Li, Mengnan [Idaho National Laboratory (INL), Idah