Search NASA⌕ Search

SEARCH · Search NASA

Results for “data system”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Making a Water Data System Responsive to Information Needs of Decision Makers

Evidence-based environmental management requires data that are sufficient, accessible, useful and used. A mismatch between data, data systems, and data needs for decision making can result in inefficient and inequitable capital investments, resource allocations, environmental protection, hazard mitigation, and quality of life. In this paper, we examine the relationship between data and decision making in environmental management, with a focus on water management. We focus on the concept of decision-driven data systems —data systems that incorporate an assessment of decision-makers' data needs into their design. The aim of the research was to examine the process of translating data into effective decision making by engaging stakeholders in the development of a water data system. Using California's legislative mandate for state agencies to integrate existing water and other environmental data as a case study, we developed and applied a participatory approach to inform data-system design and identify unmet data needs. Using workshops and focused stakeholder meetings, we developed 20 diverse use cases to assess data sources, availability, characteristics, gaps, and other attributes of data used for representative decisions. Federal and state agencies made up about 90% of the data sources, and could readily adapt to a federated data system, our recommended model for the state. The remaining 10% of more-specialized data, central to important decisions across multiple use cases, would require additional investment or incentives to achieve data consistency, interoperability, and compatibility with a federated system. Based on this assessment, we propose a typology of different types of data limitations and gaps described by stakeholders. We also propose technical, governance, and stakeholder engagement evaluation criteria to guide planning and building environmental data systems. Data-system governance involving both producers and users of data was seen as essential to achieving workable standards, stable funding, convenient data availability, resilience to institutional change, and long-term buy-in by stakeholders. Our work provides a replicable lesson for using decision-maker and stakeholder engagement to shape the design of an environmental data system, and inform a technical design that addresses both user and producer needs.

Cantor, Alida↗

Towards Auto-Generated Data Systems

After decades of progress, database management systems (DBMSs) are now the backbones of many data applications that we interact with on a daily basis. Yet, with the emergence of new data types and hardware, building and optimizing new data systems remain as difficult as the heyday of relational databases. In this paper, we summarize our work towards automating the building and optimization of data systems. Drawing from our own experience, we further argue that any automation technique must address three aspects: user specification, code generation, and result validation. We conclude by discussing a case study using videos data processing, along with opportunities for future research towards designing data systems that are automatically generated.

Computer Science↗

Challenges and alternatives to empirical orthogonal functions for earth system data

Empirical orthogonal functions (EOFs) applied to gridded Earth system data enables users to diagnose modes of variability with relative ease. Yet, many challenges to interpretation exist such that they must be used with awareness and intention when applied to gridded climate data, especially with large ensembles. Utilizing data from two different Earth system modelling large ensemble frameworks, the Energy Exoscale Earth System Model and the Community Earth System Model, as well as reanalysis data, common EOF pitfalls are summarized and discussed. Challenges include erroneous mode swapping, sign flipping, and the temporal variability of the centers of action. For modes of variability with similar contribution to variance, mode swapping is not uncommon. Sign flipping can occur with almost any mode where the pattern is correct, but the sign is arbitrary. Although the variability of the center of action is not necessarily problematic, it potentially complicates interpretation over multi-century timescales. A wide variety of alternative methods to EOFs exist, but fitness-for-purpose must be evaluated. Additionally, illustrations of alternative methods and examples of proper use are provided. Alternative methods fit into three categories: EOF variants, linear methods, and multilinear methods.

54 ENVIRONMENTAL SCIENCES↗

Reevaluating Contour Visualizations for Power Systems Data

Effective visual analytics tools are needed now more than ever as emerging energy systems data and models are rapidly growing in scale and complexity. Here we examined the suitability of colored contour maps to visually represent bus values in two different power system models: a dense 24k-bus distribution system and a 240-bus transmission system. In a quantitative analysis, we found that contour maps misrepresent power systems data, changing the statistical dispersion of the bus values, including the loss of extreme values. In a controlled empirical study with thirty professional power system research engineers, we found that these distortions significantly impact excursion identification tasks. Additionally, the engineers were less confident in their assessments using contour-based visualizations than glyph-based visualizations.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Los Alamos Enterprise Analysis and Data System (LEADS)

The Los Alamos Enterprise Analysis and Data System (LEADS) is an integrated database system and analysis application used to examine the Laboratory’s assets and equities. The data mapping framework is structured to follow laboratory organizations, programs, geographical technical areas, and other functional groupings, such as facilities and equipment. LEADS application goal is to quantify the size and information related to the current condition of assets and equities. The data and inter-dependent relationships allow for insights and further analyses to be conducted of current program-of-records, the laboratory agenda, and “what-if” scenario studies. The understanding and the outputs of this application will support understanding of the laboratory’s current state, planning, and decision-making processes.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Open‐source photovoltaic model pipeline validation against well‐characterized system data

Abstract All freely available plane‐of‐array (POA) transposition models and photovoltaic (PV) temperature and performance models in pvlib‐python and pvpltools‐python were examined against multiyear field data from Albuquerque, New Mexico. The data include different PV systems composed of crystalline silicon modules that vary in cell type, module construction, and materials. These systems have been characterized via IEC 61853‐1 and 61853‐2 testing, and the input data for each model were sourced from these system‐specific test results, rather than considering any generic input data (e.g., manufacturer's specification [spec] sheets or generic Panneau Solaire [PAN] files). Six POA transposition models, 7 temperature models, and 12 performance models are included in this comparative analysis. These freely available models were proven effective across many different types of technologies. The POA transposition models exhibited average normalized mean bias errors (NMBEs) within ±3%. Most PV temperature models underestimated temperature exhibiting mean and median residuals ranging from −6.5°C to 2.7°C; all temperature models saw a reduction in root mean square error when using transient assumptions over steady state. The performance models demonstrated similar behavior with a first and third interquartile NMBEs within ±4.2% and an overall average NMBE within ±2.3%. Although differences among models were observed at different times of the day/year, this study shows that the availability of system‐specific input data is more important than model selection. For example, using spec sheet or generic PAN file data with a complex PV performance model does not guarantee a better accuracy than a simpler PV performance model that uses system‐specific data.

14 SOLAR ENERGY↗

University Data Management Pilot Utilizing the Nuclear Research Data System

Background In 2022, the Office of Science and Technology Policy (OSTP) issued a memo that significantly reshaped the landscape of access to federally funded research. The memo mandated that all taxpayer-funded research be made available to the public without delay upon publication, without an embargo period, superseding the 2013 OSTP public access policy. This public access policy promotes transparency and the democratization of knowledge, ensuring that the fruits of scientific endeavors funded by federal agencies could be immediately accessed and built upon by scientists, educators, students, and the public at large. To implement the requirements of the OSTP guidance and DOE Public Access Plan, the Office of Nuclear Energy (NE) has implemented public access plan guidance and has identified several areas where better data management practices would further expand public access to important nuclear energy related scientific data, reports, and other technical products. Significant NE supported efforts are already underway for data management and public access to important nuclear energy related data.1 2 To address gaps in data management practices, and improve retention and accessibility of data, NE is actively exploring enhanced data management options utilizing its high-performance computing resources administered by its Nuclear Scientific User Facility Program. A newly piloted system, the Nuclear Research Data System (NRDS) acts as a portal for data collection and dissemination. Nuclear Energy University Program Research and Development Portfolio According to Web of Science, NEUP has produced 2,345 journal publication that have been cited more than 61,000 times3 and countless conference proceedings. These publications are publicly available through OSTI.gov and in the open literature. Additional scientific and technical products including project milestones that are not publications and NEUP project final reports are vetted through OSTI.gov and released once reviewed and approved by DOE. Since 2009, NEUP has awarded close to 1,000 different R&D projects in technical areas across the NE research programs. As of June 2023, 512 NEUP reports are publicly available on OSTI. The underlying data for projects is still held at universities, and data transfer, co-location, and dissemination has not occurred in a systematic way. NEUP data is currently accessible through myriad university-based data repositories, or through direct requests to PIs. The program identified this patchwork of repositories, or often lack of publicly available data, as a significant barrier to an organized, accessible, and comprehensive solution to sharing data with the larger nuclear energy community. Approach The goal of this pilot project is to establish a pathway to a consolidated long-term repository for NEUP project data. To accomplish this goal, the pilot strives to accomplish the following objectives: Establish data collection standards, including a standard set of required supplementary information to contextualize and support raw data files. Work with the HPC group collect and upload information and to modify the NRDS system, as needed, to support a standardized approach. Resolve potential barriers to successful roll out of an expanded data collection strategy, including modifying data management plan guidelines and establishing a document and data release process that accounts for potential intellectual property and/or export control concerns. Results Overall, the pilot was successful in collecting 8,982 raw and processes data files, 220 reports, 56 calibration files, and 5,931 other supplementary documents. Supplementary documents included experimental plans, methods, journal publications and conference proceedings, milestone reports, and final reports. Figure 2 shows the number of data sets and supplementary project information provided by each project. Projects has significantly different input, depending on experimental data produced and completeness of the datasets provided.

Data collection↗

Overview of IMPACT Data Acquisition System and Data Reduction Process

This report documents the development of the data acquisition system (DAS) and data reduction methodologies for the Irradiated Material Property Accelerated Characterization Test (IMPACT) experiment at the Advanced Test Reactor (ATR). The IMPACT experiment is designed to enable in-pile measurement of thermal conductivity in metallic nuclear fuels, specifically U-10Zr, using an instrumented thermal conductivity probe. The DAS supports both passive temperature monitoring and active thermal interrogation of the probe through controlled AC and DC excitation. Significant modifications to laboratory-scale systems were required to accommodate the higher resistance paths associated with the in-pile application. Custom electronics and relay-controlled measurement sequencing were developed to enable the measurement and sufficient power delivery to the sensing region. A reduced-order, axisymmetric thermal model based on the thermal quadrupoles method is presented to support data interpretation. This model enables efficient evaluation of transient heat transfer behavior and facilitates solution of the inverse problem required to extract thermal properties from measured signals. Multiple boundary condition formulations are discussed to address varying experimental time scales and geometries. Additionally, machine learning techniques are introduced to support data reduction and improve confidence in inverse solutions. Convolutional neural networks are applied to identify the presence of gas gaps and other evolving geometric features that significantly impact thermal response during irradiation. These efforts contribute to the broader integration of digital twin frameworks and real-time modeling capabilities within the Advanced Fuels Campaign.

46 - INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AN↗

Denudation, solute export, landscape evolution modeling, and geographic information system data for the East River watershed, Colorado, USA (2020-2024)

This data package contains geographic information system (GIS) layers and tabular datasets associated with the study of lithologic controls on denudation, solute export, carbon-scaling relationships, and transient landscape evolution in the East River watershed near Crested Butte, Colorado, USA. The package includes GIS layers used to produce the Figure 2 map, including drainage, hillshade, lithology, sample locations, and basin polygons, together with comma-separated value (CSV) tables and matching CSV data dictionaries. One group of tables reports sample-level and catchment-level information for river-sediment samples analyzed for in situ-produced cosmogenic beryllium-10 (10Be), including sample names, outlet elevations, geographic coordinates, upstream drainage area, rock-type classes, production-rate scaling scheme, analyzed nuclide, catchment-averaged denudation rates, and associated lower and upper analytical uncertainties. Sample and catchment attributes provide the basis for comparing denudation rates across intrusive, shale, sedimentary, and mixed-lithology settings. A second group of tables reports supporting information for landscape-evolution modeling and the mapped geologic framework of the study area. Included files list parameter values and definitions for the two-phase landscape-evolution simulations, summarize full-domain model erosion fluxes and topographic metrics for different simulation configurations, provide a fixed-area carbon-model scaling table, and summarize mapped geologic units within the East River study domain, including geologic code, formation name, lithologic description, mapped area, and lithologic class grouping. Model outputs and geologic summaries support interpretation of transient landscape behavior and its relation to the mapped distribution of shale, intrusive, sedimentary, and surficial units. A third group of tables reports hydrologic and hydrochemical information used to quantify dissolved export from the watershed. Included files provide site-level values for drainage area, mean annual solute export, standard error of annual export, area-normalized solute yield, and equivalent weathering rate for five East River monitoring sites, along with metadata describing the number, sampling cadence, and date range of discharge records and partial and full total dissolved solids observations used in the solute-yield analyses. The package also contains a supplementary daily ion-load time series with daily mean discharge, discharge observation counts, dissolved concentrations, and daily loads for calcium, magnesium, sodium, potassium, chloride, sulfate, nitrate, fluoride, dissolved silica, charge-balance bicarbonate, and total dissolved solids. The package contains GIS files, comma-separated value files (.csv), CSV data dictionaries, a file-level metadata table, a package-tree text file, and a readme text file.

10Be↗

Alternatives to Contour Visualizations for Power Systems Data

Electrical grids are geographical and topological structures whose voltage states are challenging to represent accurately and efficiently for visual analysis. The current common practice is to use colored contour maps, yet these can misrepresent the data. We examine the suitability of four alternative visualization methods for depicting voltage data in a geographically dense distribution system - Voronoi polygons, H3 tessellations, S2 tessellations, and a network-weighted contour map. We find that Voronoi tessellations and network-weighted contour maps more accurately represent the statistical distribution of the data than regular contour maps.

MATHEMATICS AND COMPUTING,POWER TRANSMISSION AND D↗

Alternatives to Contour Visualizations for Power Systems Data

Electrical grids are geographical and topological structures whose voltage states are challenging to represent accurately and efficiently for visual analysis. The current common practice is to use colored contour maps, yet these can misrepresent the data. We examine the suitability of four alternative visualization methods for depicting voltage data in a geographically dense distribution system-Voronoi polygons, H3 tessellations, S2 tessellations, and a network-weighted contour map. We find that Voronoi tessellations and network-weighted contour maps more accurately represent the statistical distribution of the data than regular contour maps.

MATHEMATICS AND COMPUTING↗

Tsunami Early Warning From Global Navigation Satellite System Data Using Convolutional Neural Networks

Abstract We investigate the potential of using Global Navigation Satellite System (GNSS) observations to directly forecast full tsunami waveforms in real time. We train convolutional neural networks to use less than 9 min of GNSS data to forecast the full tsunami waveforms over 6 hr at select locations, and obtain accurate forecasts on a test data set. Our training and test data consists of synthetic earthquakes and associated GNSS data generated for the Cascadia Subduction Zone using the MudPy software, and corresponding tsunami waveforms in Puget Sound computed using GeoClaw. We use the same suite of synthetic earthquakes and waveforms as in earlier work where tsunami waveforms were used for forecasting, and provide a comparison. We also explore varying the number of GNSS stations, their locations, and their observation durations.

Rim, Donsub↗

Rapid Detection of Anomalies in Battery Energy Storage System Data

Data analytics is pivotal in assessing the technical characteristics and performance of Battery Energy Storage Systems (BESS), underpinning BESS modeling, optimization, and control. However, raw datasets frequently harbor anomalies from measurement errors and equipment malfunctions, impacting BESS reliability and analysis accuracy To address the challenge, this paper presents a novel methodology for the rapid detection of anomalous charge or discharge cycles within BESS operational data, expediting the cleaning process while ensuring data integrity. We’ve collected diverse and comprehensive real-world BESS operational datasets in collaboration with the Electric Power Research Institute and multiple Washington State utilities. These datasets serve dual roles: enabling comprehensive data exploration and analysis for understanding underlying challenges and method development, while also acting as a vital validation resource, demonstrating practical effectiveness. The proposed method detects anomalies and aids in their resolution, improving system performance characterization precision. It also reveals recurring data anomaly sources, offering insights for data collection and handling enhancement. Practitioners can gain valuable insights from the identified anomalous cycles in the real-world datasets along with the investigative process for root cause analyses and essential data cleaning steps.

Crawford, Aladsair J.↗

Engineering-Scale Integrated Energy System Data Projection Demonstration via the Dynamic Energy Transport and Integration Laboratory

The objective of this study is to demonstrate and validate the Dynamic Energy Transport and Integration Laboratory (DETAIL) preliminary scaling analysis using Modelica language system-code Dymola. The DETAIL preliminary scaling analysis includes a multisystem integral scaling package between thermal-storage and hydrogen-electrolysis systems. To construct the system of scaled equations, dynamical system scaling (DSS) was applied to all governing laws and closure relations associated with the selected integral system. The existing Dymola thermal-energy distribution system (TEDS) facility and high-temperature steam electrolysis (HTSE) facility models in the Idaho National Laboratory HYBRID repository were used to simulate a test case and a corresponding scaled case for integrated system HYBRID demonstration and validation. The DSS projected data based on the test-case simulations and determined scaling ratios were generated and compared with scaled case simulations. The preliminary scaling analysis performance was evaluated, and scaling distortions were investigated based on data magnitude, sequence, and similarity. The results indicated a necessity to change the normalization method for thermal storage generating optimal operating conditions of 261 kW power and mass flow rate of 6.42 kg/s and the possibility of reselecting governing laws for hydrogen electrolysis to improve scaling predictive properties. To enhance system-scaling similarity for TEDS and HTSE, the requirement for scaling validation via physical-facility demonstration was identified.

08 HYDROGEN↗

Operation and Maintenance of PV Systems: Data Science, Analysis, and Standards

This effort improves the effectiveness and reduce uncertainty in O&M cost through four primary objectives/tasks: 1) institutionalize standards for reliability and availability reporting for large PV power plants; 2) bridge systemic O&M knowledge gaps around important topics affecting O&M; 3) characterize systemic failure modes and patterns and accelerate O&M experiential learning cycles using field data; and 4) establish a baseline understanding of UPVS O&M cost drivers. Key results of this effort include publication of IEC standards, published topical papers on O&M topics, training, and characterize field data for climate- and service-related patterns (additional details below). Integrating these results serves to reduce performance risk and facilitate improvement in the way solar projects are operated and maintained. Results are well received and two publications are among the most successful SETO publications at NREL ("Model of Operation and Maintenance Costs for Photovoltaic Systems with over 40,000 downloads and "Best Practices in Operation and Maintenance of PV Systems, 3rd Ed." with over 90,000 downloads).

14 SOLAR ENERGY↗

Ontologizing health systems data at scale: making translational discovery a reality

Common data models solve many challenges of standardizing electronic health record (EHR) data but are unable to semantically integrate all of the resources needed for deep phenotyping. Open Biological and Biomedical Ontology (OBO) Foundry ontologies provide computable representations of biological knowledge and enable the integration of heterogeneous data. However, mapping EHR data to OBO ontologies requires significant manual curation and domain expertise. We introduce OMOP2OBO, an algorithm for mapping Observational Medical Outcomes Partnership (OMOP) vocabularies to OBO ontologies. Using OMOP2OBO, we produced mappings for 92,367 conditions, 8611 drug ingredients, and 10,673 measurement results, which covered 68–99% of concepts used in clinical practice when examined across 24 hospitals. When used to phenotype rare disease patients, the mappings helped systematically identify undiagnosed patients who might benefit from genetic testing. By aligning OMOP vocabularies to OBO ontologies our algorithm presents new opportunities to advance EHR-based deep phenotyping.

96 KNOWLEDGE MANAGEMENT AND PRESERVATION↗