Search NASA⌕ Search

SEARCH · Search NASA

Results for “data usage”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 451 records · Page 25

Computer model to simulate testing at the National Transonic Facility

A computer model has been developed to simulate the processes involved in the operation of the National Transonic Facility (NTF), a large cryogenic wind tunnel at the Langley Research Center. The simulation was verified by comparing the simulated results with previously acquired data from three experimental wind tunnel test programs in the NTF. The comparisons suggest that the computer model simulates reasonably well the processes that determine the liquid nitrogen (LN2) consumption, electrical consumption, fan-on time, and the test time required to complete a test plan at the NTF. From these limited comparisons, it appears that the results from the simulation model are generally within about 10 percent of the actual NTF test results. The use of actual data acquisition times in the simulation produced better estimates of the LN2 usage, as expected. Additional comparisons are needed to refine the model constants. The model will typically produce optimistic results since the times and rates included in the model are typically the optimum values. Any deviation from the optimum values will lead to longer times or increased LN2 and electrical consumption for the proposed test plan. Computer code operating instructions and listings of sample input and output files have been included.

Mineck, Raymond E.↗

Large-scale Distribution of CH4 in the Western North Pacific: Sources and Transport from the Asian Continent

Methane (CH4) mixing ratios in the northern Pacific Basin were sampled from two aircraft during the TRACE-P mission (Transport and Chemical Evolution over the Pacific) from late February through early April 2001 using a tunable diode laser system. Described in more detail by Jacob et al., the mission was designed to characterize Asian outflow to the Pacific, determine its chemical evolution, and assess changes to the atmosphere resulting from the rapid industrialization and increased energy usage on the Asian continent. The high-resolution, high-precision data set of roughly 13,800 CH4 measurements ranged between 1602 ppbv in stratospherically influenced air and 2149 ppbv in highly polluted air. Overall, CH4 mixing ratios were highly correlated with a variety of other trace gases characteristic of a mix of anthropogenic industrial and combustion sources and were strikingly correlated with ethane (C2H6) in particular. Averages with latitude in the near-surface (0-2 km) show that CH4 was elevated well above background levels north of 15 deg N close to the Asian continent. In the central and eastern Pacific, levels of CH4 were lower as continental inputs were mixed horizontally and vertically during transport. Overall, the correlation between CH4 and other hydrocarbons such as ethane (C2H6), ethyne (C2H2), and propane (C3H8) as well as the urban/industrial tracer perchloroethene (C2Cl4), suggests that for CH4 colocated sources such as landfills, wastewater treatment, and fossil fuel use associated with urban areas dominate regional inputs at this time. Comparisons between measurements made during TRACE-P and those of PEM-West B, flown during roughly the same time of year and under a similar meteorological setting 7 years earlier, suggest that although the TRACE-P CH4 observations are higher, the changes are not significantly greater than the increases seen in background air over this time interval.

Bartlett, Karen B.↗

Object-Based Comparison of Data-Driven and Physics-Driven Satellite Estimates of Extreme Rainfall

The Global Precipitation Measurement (GPM) constellation of spaceborne sensors provides a variety of direct and indirect measurements of precipitation processes. Such observations can be employed to derive spatially and temporally consistent gridded precipitation estimates either via data-driven retrieval algorithms or by assimilation into physically based numerical weather models. We compare the data-driven Integrated Multisatellite Retrievals for GPM (IMERG) and the assimilation-enabled NASA-Unified Weather Research and Forecasting (NU-WRF) model against Stage IV reference precipitation for four major extreme rainfall events in the southeastern United States using an object-based analysis framework that decomposes gridded precipitation fields into storm objects. As an alternative to conventional ‘‘grid-by-grid analysis,’’ the object-based approach provides a promising way to diagnose spatial properties of storms, trace them through space and time, and connect their accuracy to storm types and input data sources. The evolution of two tropical cyclones are generally captured by IMERG and NU-WRF, while the less organized spatial patterns of two mesoscale convective systems pose challenges for both. NU-WRF rain rates are generally more accurate, while IMERG better captures storm location and shape. Both show higher skill in detecting large, intense storms compared to smaller, weaker storms. IMERG’s accuracy depends on the input microwave and infrared data sources; NU-WRF does not appear to exhibit this dependence. Findings highlight that an object-oriented view can provide deeper insights into satellite precipitation performance and that the satellite precipitation community should further explore the potential for ‘‘hybrid’’ data-driven and physics-driven estimates in order to make optimal usage of satellite observations.

extreme events↗

Electric Propulsion for the Psyche Mission: Development Activities and Status

NASA’s Psyche mission will launch in 2022 and begin a 3.6-year cruise to the metallic asteroid Psyche, where it will examine this unique body. The baseline spacecraft design is a hybrid of JPL’s deep-space heritage subsystems with commercial partner Maxar’s electric propulsion, power, and structure subsystems. All primary propulsion will be done with SPT-140 thrusters, which will be the first use of Hall thrusters for a NASA mission. The electric propulsion subsystem and its implementation for the Psyche mission are described here. Major testing activities have included the successful completion of subsystem integrated testing with the design modifications required for Psyche, and a series of low-power thrust repeatability tests that were performed in support of navigation analyses. Thruster performance models have been further validated with new SPT-140 flight data, and new analyses of thruster swirl torque have been performed that result in much higher values than previously estimated. Analysis of recent Maxar flight data has also provided a new understanding of in-flight propellant usage uncertainties. Subsystem integration and test activities are now underway and the status and plans are discussed.

Johnson, Ian↗

EVs@Scale Next-Gen Profiles - Fleet Utilization 2024

As part of the U.S. Department of Energy’s EVs@Scale initiative, the Next-Gen Profiles (NGP) project provides a comprehensive, data-driven analysis of electric vehicle (EV) and electric vehicle supply equipment (EVSE) operations across real-world fleet deployments. This paper presents findings from the NGP’s Fleet Utilization study, which investigates operational behavior and asset usage across seventeen EV fleets and two EVSE fleets, encompassing a wide range of vehicle types and use cases. Data collected from diverse sources—varying in format and temporal resolution—are first reformatted into a unified structure. From this harmonized dataset, a suite of rigorously defined performance metrics is calculated at an hourly cadence, enabling consistent cross-comparison of charging, routing, and other key operational behaviors. Amid rapidly increasing EV adoption and growing demands for energy-efficient fleet operations, the analysis reveals clear utilization trends—including diurnal and weekly activity cycles, differences in short versus long charging session dependencies, and route-specific energy usage patterns. These findings highlight the need for tailored infrastructure strategies and the deployment of advanced energy management systems, such as Distributed Energy Resource Management Systems (DERMS) and Site Energy Management Systems (SEMS), which can optimize charging schedules and mitigate peak loads. By leveraging anonymized, harmonized datasets and standardized metrics, this study offers critical insights into fleet behavior and performance, providing a foundation to improve operational efficiency, reduce costs, and enable the scalable deployment of electrified transportation.

Wells, Landon↗

h5bench: A unified benchmark suite for evaluating HDF5 I/O performance on pre‐exascale platforms

Summary Parallel I/O is a critical technique for moving data between compute and storage subsystems of supercomputers. With massive amounts of data produced or consumed by compute nodes, high‐performant parallel I/O is essential. I/O benchmarks play an important role in this process; however, there is a scarcity of I/O benchmarks representative of current workloads on HPC systems. Toward creating representative I/O kernels from real‐world applications, we have created h5bench , a set of I/O kernels that exercise hierarchical data format version 5 (HDF5) I/O on parallel file systems in numerous dimensions. Our focus on HDF5 is due to the parallel I/O library's heavy usage in various scientific applications running on supercomputing systems. The various tests benchmarked in the h5bench suite include I/O operations (read and write), data locality (arrays of basic data types and arrays of structures), array dimensionality (one‐dimensional arrays, two‐dimensional meshes, three‐dimensional cubes), I/O modes (synchronous and asynchronous). In this paper, we present the observed performance of h5bench executed along several of these dimensions on existing supercomputers (Cori and Summit) and pre‐exascale platforms (Perlmutter, Theta, and Polaris). h5bench measurements can be used to identify performance bottlenecks and their root causes and evaluate I/O optimizations. As the I/O patterns of h5bench are diverse and capture the I/O behaviors of various HPC applications, this study will be helpful to the broader supercomputing and I/O community.

97 MATHEMATICS AND COMPUTING↗

New U.S. Data Tools are Playing a Crucial Role in Decarbonizing Buildings at Speed, Scale, and Low Cost

Preparing buildings for retrofits traditionally requires expensive on-site audits or timeintensive simulation models. As a result, the majority of buildings fail to pursue cost-saving retrofits. To address these barriers, the U.S. Department of Energy (DOE) has introduced the Building Efficiency Targeting Tool for Energy Retrofits (BETTER)-a new, free, on-line tool that utilizes a data-driven analytical engine and user-friendly web interface to automatically analyze a building's monthly energy usage in response to weather conditions. The tool benchmarks a building's electric and fossil energy usage against peers; estimates energy, cost, and emissions reductions at the building and portfolio levels; recommends energy efficiency measures; and prioritizes buildings for net-zero energy retrofits. Thanks to interoperability with the DOE's Standard Energy Efficiency Data (SEED) platform, BETTER is supporting U.S. jurisdictions to prepare buildings for retrofit at speed, scale, and low cost to comply with energy policies. This paper discusses the use of BETTER and SEED by one of the branches of the California state government to streamline a retrofit program across 455 public non-residential buildings to align with state goals to reduce greenhouse gas emissions. It describes the organization's challenge to reduce energy consumption across a geographically diverse, aging portfolio; explores how BETTER and SEED improved workflow efficiency; presents preliminary results, including avoiding audit costs of $3.28 million and developing the groundwork for retrofit projects estimated to prevent emission of 2,271 t CO2e annually; and provides guidance for other jurisdictions seeking similar results.

BETTER↗

New U.S. Data Tools are Playing a Crucial Role in Decarbonizing Buildings at Speed, Scale, and Low Cost

Preparing buildings for retrofits traditionally requires expensive on-site audits or time- intensive simulation models. As a result, the majority of buildings fail to pursue cost-saving retrofits. To address these barriers, the U.S. Department of Energy (DOE) has introduced the Building Efficiency Targeting Tool for Energy Retrofits (BETTER)—a new, free, on-line tool that utilizes a data-driven analytical engine and user-friendly web interface to automatically analyze a building’s monthly energy usage in response to weather conditions. The tool benchmarks a building’s electric and fossil energy usage against peers; estimates energy, cost, and emissions reductions at the building and portfolio levels; recommends energy efficiency measures; and prioritizes buildings for net-zero energy retrofits. Thanks to interoperability with the DOE’s Standard Energy Efficiency Data (SEED) platform, BETTER is supporting U.S. jurisdictions to prepare buildings for retrofit at speed, scale, and low cost to comply with energy policies. This paper discusses the use of BETTER and SEED by one of the branches of the California state government to streamline a retrofit program across 455 public non-residential buildings to align with state goals to reduce greenhouse gas emissions. It describes the organization’s challenge to reduce energy consumption across a geographically diverse, aging portfolio; explores how BETTER and SEED improved workflow efficiency; presents preliminary results, including avoiding audit costs of $3.28 million and developing the groundwork for retrofit projects estimated to prevent emission of 2,271 t CO2e annually; and provides guidance for other jurisdictions seeking similar results.

Li, han↗

ComStock Measure Documentation: Variable-Speed Pumps

Building on the 3-year End-Use Load Profiles project to calibrate and validate the U.S. Department of Energy's ResStock and ComStock models, this work produces national data sets that enable cities, states, utilities, and other stakeholders to answer a broad range of questions regarding their commercial building stock. ComStock is a highly granular, bottom-up model that uses various data sources, statistical sampling methods, and advanced building energy simulations to estimate the annual subhourly energy consumption of the commercial building stock across the United States. The "baseline" model intends to represent the U.S. commercial building stock as it existed in 2018. The methodology of the baseline model is discussed in the ComStock Reference Documentation. The goal of this work is to develop energy efficiency and demand flexibility measures that cover market-ready technologies and study their mass adoption impact on the baseline building stock. "Measures" refers to various "what-if" scenarios that can be applied to buildings. The results for the baseline and measure scenario simulations are published in public data sets that provide insights into building stock characteristics, operational behaviors, utility bill impacts, and annual and sub-hourly energy usage by fuel type and end use. This report describes the modeling methodology for a single ComStock measure scenario - variable speed pumps - and briefly introduces key results. The full public data set can be accessed on the ComStock data lake or via the Data Viewer at comstock.nlr.gov. The public data set enables users to create custom aggregations of results for their use case (e.g., filter to a specific county or building type).

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

PNNL-Predictive-Phenomics/ProteoMeter

ProteoMeter is a Python package that assists in the statistical analysis of global proteomics, protein post-translation modification (PTM), and limited proteolysis (LiP) data. It contains batch correction, normalization, and statistical testing methods, as well as functions that "roll up" peptide-level data to the single-site level. It has a robust user configuration system, allowing it to flexibly integrate different types of experiment designs. For basic usage, a simple configuration file provides the essential functionality. Advanced users have access to the entire statistical pipeline for fine-tuning analyses. Processed data is easily exported to many common spreadsheet and data-frame formats.

Rozum, Jordan [Pacific Northwest National Lab]↗

From roads to roofs: How urban and rural mobility influence building energy consumption

In this article, understanding the relationship between travel behavior and building energy use at an urban scale is crucial for developing effective energy management strategies. Mobility patterns significantly impact building occupancy, which in turn affects energy consumption. However, existing methods often focus on individual buildings, whereas geographical influences on energy usage are not adequately examined. This study addresses this gap by using transportation origin-destination (OD) data to estimate building occupancy and energy. The proposed method assigns OD trips from census block groups to the building level, incorporating building, travel survey, and census data to derive building occupancy profiles. This method was applied to urban and rural areas with 4062 buildings in 70 census block groups. We found that the OD-informed occupancy profile exhibits smoother energy consumption patterns compared with that of Department of Energy reference occupancy profiles. Our analysis reveals distinct building energy consumption patterns among groups with long and short commutes, emphasizing the effect of commute times and work schedules on residential energy usage. This framework is useful for practitioners in transportation agencies and utility companies, enabling the estimation of building energy based on mobility patterns. Overall, this study shows the potential of integrating transportation and building energy data to inform cross-sector energy management strategies.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

The critical importance of software for HEP

Particle physics has an ambitious and broad global experimental programme for the coming decades. Large investments in building new facilities are already underway or under consideration. Scaling the present processing power and data storage needs by the foreseen increase in data rates in the next decade for HL-LHC is not sustainable within the current budgets. As a result, a more efficient usage of computing resources is required in order to realise the physics potential of future experiments. Software and computing are an integral part of experimental design, trigger and data acquisition, simulation, reconstruction, and analysis, as well as related theoretical predictions. A significant investment in computing and software is therefore critical. Advances in software and computing, including artificial intelligence (AI) and machine learning (ML), will be key for solving these challenges. Making better use of new processing hardware such as graphical processing units (GPUs) or ARM chips is a growing trend. This forms part of a computing solution that makes efficient use of facilities and contributes to the reduction of the environmental footprint of HEP computing. The HEP community already provided a roadmap for software and computing for the last EPPSU, and this paper updates that, with a focus on the most resource critical parts of our data processing chain.

97 MATHEMATICS AND COMPUTING↗

Same Data, Different Audiences: Using Personas to Scope a Supercomputing Job Queue Visualization

Domain-specific visualizations sometimes focus on narrow, albeit important, tasks for one group of users. This focus limits the utility of a visualization to other groups working with the same data. While tasks elicited from other groups can present a design pitfall if not disambiguated, they also present a design opportunity—namely, the development of visualizations that support multiple groups. This development choice presents a trade-off of broadening the scope but limiting support for the more narrow tasks of any one group, which in some cases can enhance the overall utility of the visualization. We investigate this scenario through a design study where we develop Guidepost, a notebook-embedded visualization of data that helps scientists assess compute wait times, machine learning researchers understand prediction accuracy, and system maintainers analyze usage trends. We adapt the use of personas for visualization design from existing literature in the HCI and design domains, applying them to categorize tasks based on their uniqueness across stakeholder personas. Under this model, tasks shared between all groups should be supported by interactive visualizations and tasks unique to each group can be deferred to scripting with notebook-embedded visualization design. We evaluate our visualization through real-world case studies and a task-focused evaluation with nine participants. We observe that together, Guidepost's visual encodings, interactions, and export capabilities support the tasks of our differing personas.

97 MATHEMATICS AND COMPUTING↗

IGES, a key interface specification for CAD/CAM systems integration

The Initial Graphics Exchange Specification (IGES) program has focused the efforts of 52 companies on the development and documentation of a means of graphics data base exchange among present day CAD/CAM systems. The project's brief history has seen the evolution of the Specification into preliminary industrial usage marked by public demonstrations of vendor capability, mandatory requests in procurement actions, and a formalization into an American National Standard in September 1981. Recent events have demonstrated intersystem data exchange among seven vendor systems with a total of 30 vendors committing to offer IGES capability. A full range of documentation supports the IGES project and the recently approved IGES Version 2.0 of the Specification.

Smith, B. M.↗

ComStock Measure Documentation: High-Efficiency Rooftop Unit

Building on the 3-year End-Use Load Profiles project to calibrate and validate the U.S. Department of Energy's ResStock and ComStock models, this work produces national data sets that enable cities, states, utilities, and other stakeholders to answer a broad range of questions regarding their commercial building stock. ComStock is a highly granular, bottom-up model that uses various data sources, statistical sampling methods, and advanced building energy simulations to estimate the annual subhourly energy consumption of the commercial building stock across the United States. The "baseline" model intends to represent the U.S. commercial building stock as it existed in 2018. The methodology of the baseline model is discussed in the ComStock Reference Documentation. The goal of this work is to develop energy efficiency and demand flexibility measures that cover market-ready technologies and study their mass adoption impact on the baseline building stock. "Measures" refers to various "what-if" scenarios that can be applied to buildings. The results for the baseline and measure scenario simulations are published in public data sets that provide insights into building stock characteristics, operational behaviors, utility bill impacts, and annual and sub-hourly energy usage by fuel type and end use. This report describes the modeling methodology for a single ComStock measure scenario - high-efficiency rooftop unit (RTU) - and briefly introduces key results. The full public data set can be accessed on the Comstock data lake or via the Data Viewer at comstock.nlr.gov. The public data set enables users to create custom aggregations of results for their use case (e.g., filter to a specific county or building type).

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

MODIS Technical Report Series. Volume 4: MODIS data access user's guide: Scan cube format

The software described in this document provides I/O functions to be used with Moderate Resolution Spectroradiometer (MODIS) level 1 and 2 data, and could be easily extended to other data sources. This data is in a scan cube data format: a 3-dimensional ragged array containing multiple bands which have resolutions ranging from 250 to 1000 meters. The complexity of the data structure is handled internally by the library. The I/O calls allow the user to access any pixel in any band through 'C' structure syntax. The high MODIS data volume (approaching half a terabyte per day) has been a driving factor in the library design. To avoid recopying data for user access, all I/O is performed through dynamic 'C' pointer manipulation. This manual contains background material on MODIS, several coding examples of library usage, in-depth discussions of each function, reference 'man' type pages, and several appendices with details of the included files used to customize a user's data product for use with the library.

Kalb, Virginia L.↗

Development of a Distribution Optimal Power Flow Federate for Open-Source OEDI-SI Platform

Increasing numbers of distributed generators in the electric power distribution networks require developing a control strategy to optimize solutions in real time. Linearized optimal distribution flow development has seen growth and acceptance in the distribution systems literature for efficiently modeling the \glspl{opf} for distribution systems. This paper examines the implementation and integration procedure for linearized optimal distribution flow federate to \gls{oedisi} platform. Specifically, we discuss i) the usage of the \gls{oedisi} platform, ii) obtaining a tractable solution using developed \gls{opf} federate, and iii) validation of solutions and bench-marking the \gls{oedisi} platform with developed \gls{opf} federate using OpenDSS. In brief, we demonstrate how a general linearized optimal distribution flow federate can be developed and integrated with a co-simulation environment to mimic real-world examples. The efficacy of the proposed method is demonstrated using the IEEE 123-bus test system under different scenarios to obtain a tractable solution and compare its results.

Sadnan, Rabayet↗

On the Abuse and Detection of Polyglot Files

A polyglot is a file that is valid in two or more formats. Polyglot files pose a problem for file-upload and generative AI web interfaces that rely on format identification to determine how to securely handle incoming files. In this work we found that existing file-format and embedded-file detection tools, even those developed specifically for polyglot files, fail to reliably detect polyglot files used in the wild. To address this issue, we studied the use of polyglot files by malicious actors in the wild, finding 30 polyglot samples and 15 attack chains that leveraged polyglot files. Using knowledge from our survey of polyglot usage in the wild---the first of its kind---we created a novel data set based on adversary techniques. We then trained a machine learning detection solution, PolyConv, using this data set. PolyConv achieves a precision-recall area-under-curve score of 0.999 with an F1 score of 99.20% for polyglot detection and 99.47% for file-format identification, significantly outperforming all other tools tested. We developed a content disarmament and reconstruction tool, ImSan, that successfully sanitized 100% of the tested image-based polyglots, which were the most common type found via the survey. Our work provides concrete tools and suggestions to enable defenders to better defend themselves against polyglot files, as well as directions for future work to create more robust file specifications and methods of disarmament.

Oesch, T [ORNL] (ORCID:0000000269091022)↗