Search NASA⌕ Search

SEARCH · Search NASA

Results for “data sharing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 307 records · Page 17

If We Build Them, They Will Run: Automated HPC Apps Deployment and Profiling with eBPF in Cloud

The high performance computing (HPC) community is in a period of transition. The rise of AI/ML coupled with a changing landscape of resources deems portability a new metric of performance, and methods to move between on-premises and cloud environments and assess compatibility are paramount. Here we design and test a strategy for bridging the gap between traditional HPC and Kubernetes environments – first containerizing applications, providing automated orchestration to run studies, and packaging the setup with automated means to assess performance using low overhead eXtended Berkeley Packet Filter (eBPF) programs. We first assess different designs for eBPF collection, demonstrating a tradeoff between number of programs deployed on a node and overhead added. We develop 5 low overhead eBPF programs that combine with streaming ML models to assess CPU, futex, TCP, shared memory, and file access across four different builds of an HPC application for CPU and GPU. We use eBPF data to generate insights into the possible underlying etiology of scaling issues. We then assess compatibility of a well-known benchmark, HPCG, across matrices of micro-architectures and optimization levels (217 containers across 24 instance types and over 7500 runs). We provide to the community 30 applications to deploy in our automated setup and perform a scaling study from 4 to a maximum of 256 nodes for both CPU and GPU applications. Finally, we use our gained knowledge about performance to generate compatibility artifacts that are used by a newly developed Kubernetes controller to intelligently select instance type based on optimizing a figure of merit. Along with insights to scaling in this environment with a collection of applications and templates to work from, we provide an overall strategy for approaching HPC application deployment and image selection based on compatibility in cloud.

Computer science↗

The ABCs of On-Demand Transit (ODT)

On-demand mobility - also referred to as on-demand transit (ODT) - is a form of public mobility that is flexible with respect to where and when service is provided, and ODT deployments have increased significantly in recent years. Transit agencies are becoming increasingly interested in ODT, and due to differing definitions and various service design and business model options, it can be difficult to learn about the emerging industry. This work provides an overview of the definitions of ODT, recent trends internationally and in the U.S., ODT's benefits and challenges (particularly compared to fixed-route transit), three primary service design options, system costs and funding considerations, and a metrics framework for evaluating ODT systems to ensure continued successful performance. Identified benefits include increased service areas, short ride and wait times, increased user flexibility, potential to reduce energy consumption and emissions through shared trips and smaller, right-sized vehicles, increased safety and comfort through door-to-door service, and rich data streams including granular spatio-temporal data that can be analyzed to continuously improve the service. Challenges include scaling ODT service up as small increases in ridership require additional supply to keep service quality high, serving peak times including keeping low wait times, the lack of fixed schedule being challenging for commuters, integrating ODT services with nearby transit systems, and equity for riders without smartphones who cannot track the vehicle in a mobile app. Finally, an overview of seven ODT case studies (in Texas, Missouri, New York, and Ontario, Canada) performed by NREL and related analysis of travel time, energy and emissions, costs, and equity are presented. Initial key findings include: ODT can be cost- and energy-effective compared to fixed-route transit, ODT serves more people than other transit options, and ODT system deployments can be followed by rapid growth.

24 POWER TRANSMISSION AND DISTRIBUTION↗

U.S. Distribution Transformer Demand Phase III - Key Drivers and Managing Demand [Slides]

This presentation demonstrates a significant analysis, on forecasting the demand for distribution transformers. The analysis is conducted for the United States, estimating the initial in-service capacity of these assets, and forecasting demand for these assets through 2050 with several sensitivities conducted. It examines not only demand for these assets, but importantly, how utility planning practices can impact the demand for theses assets in time. Under load growth scenarios, whether utilities practice like-for-like replacement strategies as failures occur, or whether they practice proactive up-sizing, anticipating electric load growth, can have major impacts on future demand. It also examines several other growth factors, such as the increasing demand for step-up transformers, which share many of the same characteristics as distribution transformers, and the demand for specific transformers for large project growth from data centers.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Detector Characterization for SuperCDMS Commissioning

Super Cryogenic Dark Matter Search (SuperCDMS) SNOLAB is a next generation direct detection experiment search ing for low mass dark matter using cryogenic germanium and silicon detectors operated at millikelvin temperatures. As the experiment begins its first commissioning data taking, establishing that the detectors respond to energy deposits in a stable, predictable way is a prerequisite for any future physics analysis. This work presents a study of detector stability for four SuperCDMS SNOLAB detectors, det 7 and det 15 (germanium), and det 11 and det 14 (silicon), using two data sets taken during early commissioning: dedicated Barium-133 calibration runs, which provide a known gamma ray energy reference at 356 keV, and low background runs, which record whatever background radiation the detectors see with no external source present. The Ba-133 data do not show a distinct, well localized line at the expected energy, and the low background data show a baseline that drifts and oscillates over time rather than remaining flat. This baseline instability appears consistently across multiple channels rather than being confined to one, suggesting a shared, detector wide cause rather than a single faulty channel. Together, these observations point to the detectors’ cryogenic support system as the likely source of the instability, since small temperature fluctuations introduced during normal operation of the cooling system could plausibly couple into the exquisitely temperature sensitive detectors. These results inform the ongoing commissioning effort by narrowing down where instability in the current data is originating from.

O'Hanlon, Viktoria M. [Skidmore Coll.; Fermilab]↗

Detector Characterization for SuperCDMS Commissioning

Super Cryogenic Dark Matter Search (SuperCDMS) SNOLAB is a next generation direct detection experiment search ing for low mass dark matter using cryogenic germanium and silicon detectors operated at millikelvin temperatures. As the experiment begins its first commissioning data taking, establishing that the detectors respond to energy deposits in a stable, predictable way is a prerequisite for any future physics analysis. This work presents a study of detector stability for four SuperCDMS SNOLAB detectors, det 7 and det 15 (germanium), and det 11 and det 14 (silicon), using two data sets taken during early commissioning: dedicated Barium-133 calibration runs, which provide a known gamma ray energy reference at 356 keV, and low background runs, which record whatever background radiation the detectors see with no external source present. The Ba-133 data do not show a distinct, well localized line at the expected energy, and the low background data show a baseline that drifts and oscillates over time rather than remaining flat. This baseline instability appears consistently across multiple channels rather than being confined to one, suggesting a shared, detector wide cause rather than a single faulty channel. Together, these observations point to the detectors’ cryogenic support system as the likely source of the instability, since small temperature fluctuations introduced during normal operation of the cooling system could plausibly couple into the exquisitely temperature sensitive detectors. These results inform the ongoing commissioning effort by narrowing down where instability in the current data is originating from.

O'Hanlon, Viktoria M. [Skidmore Coll.; Fermilab]↗

Applying the FAIR Principles to computational workflows

Recent trends within computational and data sciences show an increasing recognition and adoption of computational workflows as tools for productivity and reproducibility that also democratize access to platforms and processing know-how. As digital objects to be shared, discovered, and reused, computational workflows benefit from the FAIR principles, which stand for Findable, Accessible, Interoperable, and Reusable. The Workflows Community Initiative’s FAIR Workflows Working Group (WCI-FW), a global and open community of researchers and developers working with computational workflows across disciplines and domains, has systematically addressed the application of both FAIR data and software principles to computational workflows. We present recommendations with commentary that reflects our discussions and justifies our choices and adaptations. These are offered to workflow users and authors, workflow management system developers, and providers of workflow services as guidelines for adoption and fodder for discussion. The FAIR recommendations for workflows that we propose in this paper will maximize their value as research assets and facilitate their adoption by the wider community.

97 MATHEMATICS AND COMPUTING↗

MINERvA s Open Data Product: A First for Neutrino Data Preservation

Access to information on neutrino nucleus interactions is critical to the success of all neutrino oscillation experiments. MINERvA's rich dataset covers a range of energies and nuclei unique amongst experiments, and as such is critical to the community in building the important shared knowledge needed to unravel the mysteries of the neutrino. In particular, its dataset provides the greatest statistical coverage in in the range of neutrino energies pertinent for DUNE until DUNE's near detector begins operation. Historically, such significant datasets in neutrino physics have been preserved primarily through their published results. While meaningful and useful, this limits the ability to explore the data to its fullest extent as new perspectives continue to form. MINERvA has undertaken a major effort to break this trend and preserve its data in a format to be as analyzable as possible from outside the collaboration. This has culminated in the officially-released MINERvA Open Data Product for the community to take advantage of and utilize. Maintaining direct access to the dataset in an analyzable form will allow new insights to continue to be extracted indefinitely. This talk will cover the contents of this product, the information included (and excluded), the tools provided to utilize the product effectively, the support MINERvA intends to provide in its use, and some lessons learned through the process.

Last, David [Rochester U.] (ORCID:0000000245147183↗

Energy-efficient multimodal mobility networks in transportation digital twins: Strategies and optimization

The study proposes a comprehensive Transportation Mobility (TransitMo) framework covering conceptual design, model formulation, optimization, simulation, and impact analysis of the transportation mobility system. TransitMo is composed of a transportation digital twin developed in Simulation of Urban MObility (SUMO) and an Intelligent Traffic Management and Control Center (ITMCC) that identifies the best ways to improve the movement of people within urban areas using various modes of transportation. This study encompasses advanced modeling techniques, algorithms, and strategic testing to optimize energy efficiency and mobility in a multimodal shared mobility network. TransitMo’s practical applications are exemplified through a city-scaled simulation network in Chattanooga, TN, employing demographic data to analyze historical traffic patterns and forecast future demands. Central to this methodology are three models: the User Preference Model (UP), the Energy Consumption Model (EC), and the System Optimization Model (SO). These models work in concert to iteratively devise the optimal travel incentives and minimize the total system cost in a real-time manner. In conclusion, test results verified that the proposed adaptive incentive program and optimized bus scheduling can improve network performance by increasing public transit ridership.

42 ENGINEERING↗

cclib 2.0: An updated architecture for interoperable computational chemistry

Interoperability in computational chemistry is elusive, impeded by the independent development of software packages and idiosyncratic nature of their output files. The cclib library was introduced in 2006 as an attempt to improve this situation by providing a consistent interface to the results of various quantum chemistry programs. The shared API across programs enabled by cclib has allowed users to focus on results as opposed to output and to combine data from multiple programs or develop generic downstream tools. Initial development, however, did not anticipate the rapid progress of computational capabilities, novel methods, and new programs; nor did it foresee the growing need for customizability. Here, we recount this history and present cclib 2, focused on extensibility and modularity. We also introduce recent design pivots—the formalization of cclib’s intermediate data representation as a tree-based structure, a new combinator-based parser organization, and parsed chemical properties as extensible objects.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

The secondary metabolism collaboratory: a database and web discussion portal for secondary metabolite biosynthetic gene clusters

Secondary metabolites are small molecules produced by all corners of life, often with specialized bioactive functions with clinical and environmental relevance. Secondary metabolite biosynthetic gene clusters (BGCs) can often be identified within DNA sequences by various sequence similarity tools, but determining the exact functions of genes in the pathway and predicting their chemical products can often only be done by careful, manual comparative analysis. To facilitate this, we report the first release of the secondary metabolism collaboratory (SMC), which aims to provide a comprehensive, tool-agnostic repository of BGC sequence data drawn from all publicly available and user-submitted bacterial and archaeal genome and contig sources. On the website, users are provided a searchable catalog of putative BGCs identified from each source, along with visualizations of gene and domain annotations derived from multiple sequence analysis tools. SMC’s data is also available through publicly-accessible application programming interface (API) endpoints to facilitate programmatic access. Users are encouraged to share their findings (and search for others’) through comment posts on BGC and source pages. At the time of writing, SMC is the largest repository of BGC information, holding 13.1M BGC regions from 1.3M source sequences and growing, and can be found at https://smc.jgi.doe.gov.

59 BASIC BIOLOGICAL SCIENCES↗

MIBiG 4.0: advancing biosynthetic gene cluster curation through global collaboration

Specialized or secondary metabolites are small molecules of biological origin, often showing potent biological activities with applications in agriculture, engineering and medicine. Usually, the biosynthesis of these natural products is governed by sets of co-regulated and physically clustered genes known as biosynthetic gene clusters (BGCs). To share information about BGCs in a standardized and machine-readable way, the Minimum Information about a Biosynthetic Gene cluster (MIBiG) data standard and repository was initiated in 2015. Since its conception, MIBiG has been regularly updated to expand data coverage and remain up to date with innovations in natural product research. Here, we describe MIBiG version 4.0, an extensive update to the data repository and the underlying data standard. In a massive community annotation effort, 267 contributors performed 8304 edits, creating 557 new entries and modifying 590 existing entries, resulting in a new total of 3059 curated entries in MIBiG. Particular attention was paid to ensuring high data quality, with automated data validation using a newly developed custom submission portal prototype, paired with a novel peer-reviewing model. MIBiG 4.0 also takes steps towards a rolling release model and a broader involvement of the scientific community. MIBiG 4.0 is accessible online at https://mibig.secondarymetabolites.org/.

59 BASIC BIOLOGICAL SCIENCES↗

Novel Approach to PV Inverter Modeling and Simulation Leveraging Experiments, Learning Based Modeling and Co-Simulation

Photovoltaic (PV) inverter manufacturers use custom, proprietary control approaches and topologies in their inverter design. The proprietary nature of these approaches makes it challenging to share electromagnetic transients (EMT) domain models for system studies. This research work presents an approach to develop EMT models from experimental data. We use novel approach in experimental design, high fidelity data collection, use of learning-based modeling, and co-simulation to reduce the time taken to develop an EMT model for an inverter under test (IUT). We used a 20 kW off-the-shelf grid following PV inverter and subjected the inverter to controlled tests. The tests include voltage and frequency step changes, as well as solar irradiance variations. The recorded high frequency data were used to train a neural network model representing the dynamic behavior of the IUT. The model was subsequently imported into an EMT tool using co-simulation techniques, and thus completing the modeling effort.

black box inverter modeling↗

Application of Flex-QA Arrays in HTS Magnet Testing

Flexible PCB quench antennas have been very useful in providing high-quality high-resolution data in low temperature superconducting magnet tests. Similar multi-sensor arrays have been employed recently to cover a high temperature superconductor magnet tested at FNAL. In the present work, data taking conditions and magnet features to support the analysis framework are discussed. Then observations made during complete magnet powering cycles are described and analysis of quench antenna data are presented. Based on results, improvements to instrumentation and data taking are debated. Views on the future of flexible quench antenna sensors for HTS magnet diagnostics and operational support are shared.

43 PARTICLE ACCELERATORS↗

Application of Flex-QA Arrays in HTS Magnet Testing

Flexible PCB quench antennas have been very useful in providing high-quality high-resolution data in low temperature superconducting magnet tests. Similar multi-sensor arrays have been employed recently to cover a high temperature superconductor magnet tested at FNAL. In the present work, data taking conditions and magnet features to support the analysis framework are discussed. Then observations made during complete magnet powering cycles are described and analysis of quench antenna data are presented. Based on results, improvements to instrumentation and data taking are debated. Views on the future of flexible quench antenna sensors for HTS magnet diagnostics and operational support are shared.

Stoynev, S.↗

Deploying and Operating CephFS for Scientific Applications at Fermilab

Fermilab has been running a Ceph cluster in production for several years to support high-throughput scientific computing. Our primary use case is CephFS, which serves interactive data analysis workloads, with growing interest in using RGW for scalable object storage of scientific datasets. In this talk, we'll share lessons learned from successfully deploying and maintaining our Ceph cluster with cephadm, including challenges faced, performance tuning, and operational practices. We'll also present custom tools we've developed to streamline monitoring and management and discuss how Ceph fits into our broader storage architecture for large-scale scientific research.

Peisker, Alison [Fermilab]↗

Cloud-Based Demonstration of the Eastern Interconnection Situational Awareness Monitoring System (ESAMS)

This report describes a cloud-based implementation and field demonstration of the Eastern Interconnection Situational Awareness and Monitoring System (ESAMS). ESAMS was developed to support the detection and source localization of forced oscillations using synchrophasor measurements from tie-lines connecting areas served by different reliability coordinators (RCs), so that RCs could better coordinate their response to wide-area events. A previous effort had identified deployment barriers associated with hosting shared situational awareness tools at a single RC. To address these barriers, ESAMS was migrated to Amazon Web Services and evaluated in a six-month field demonstration. ISO New England (ISO-NE) and PJM streamed data to the platform using AWS Direct Connect and a site-to-site VPN, respectively. The resulting multi-utility measurement footprint enabled regional source localization across major portions of the U.S. Eastern Interconnection and supported routine identification of oscillation events. During the final three months of the trial, 24 events above 2 MW/MVAR were detected. The largest detected oscillation approached a 25 MW peak-to-peak amplitude, and the longest persisted intermittently for more than 11 hours. The demonstration also assessed operational considerations—including data transfer volumes, end-to-end latency, and cloud computing costs—and found that network and compute requirements were modest relative to typical cloud capabilities while providing performance comparable to prior on-premises deployments. Overall, the results indicate that cloud hosting can provide a practical path to shared interconnection-wide oscillation monitoring. The cloud ESAMS demonstration establishes a foundation for broader utility participation and for building future wide-area analytics that leverage measurements across organizational boundaries.

Follum, James D.↗

Inverter Model Validation and Calibration Using Phasor Measurement Unit Data

As the penetration of inverter-based renewable energy resources increases in the power grid, especially at the distribution and microgrid levels, the need to accurately represent them in planning studies increases as well. However, due to the lack of well-established standard procedures, and vendor reluctance towards the detailed sharing of proprietary models, automated dynamic model validation and parameter calibration tools for inverter based resources (IBRs) remain scarce. This work presents a model validation and parameter calibration platform for representing IBRs with generic phasor-domain models. Phasor measurements of power system events are used for continuous validation using the data playback method, and model parameters are re-calibrated if a significant mismatch between measurements and model response is observed. Unique features of the proposed platform include- (a) an iterative Bayesian optimization approach towards parameter calibration to address a possible mismatch between the structures of generic models implemented in simulation softwares and actual commercial inverters, (b) error metrics designed to account for a possible mismatch between the time resolution of simulation and measurements, and (c) analysis of the measurement-simulation mismatch to provide guidance to engineering personnel regarding model shortcomings. The performance of the platform has been illustrated using both simulated data and field measurements to validate/calibrate inverter models in GridLAB-D.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Local structure of zinc–indium–tin oxide films via grazing-incidence x-ray pair-distribution functions and theoretical methods

A detailed experimental and theoretical study on the local (r ≤ 4.5 Å) atomic structure of amorphous and crystalline zinc–indium–tin oxide (ZITO) thin films using grazing-incidence x-ray Pair-Distribution Functions (PDFs), ab initio Molecular Dynamics (MD), and Empirical Potential Structure Refinement (EPSR) Monte Carlo simulations is presented. High-energy synchrotron x rays, a two-dimensional detector, and different incident angles were used to probe the depth uniformity of five (ZnO) 0.15 (In 2 O 3 ) 0.70 (SnO 2 ) 0.15 films that were deposited via pulsed-laser deposition at growth temperatures (T G ) ranging from 25 to 300 °C. Films deposited at T G ≤ 150 °C were amorphous. The partially crystalline (T G = 200 °C) and fully crystalline (T G = 300 °C) films were highly textured. Both crystalline and amorphous structures were investigated using ab initio MD and EPSR Monte Carlo simulations. The density of the amorphous films determined from the experimental data agreed with MD calculations. Coordination numbers, bond lengths, and distortion for metal–oxygen and for both the edge- and corner-shared In–metal shells up to 4.5 Å obtained from PDF analysis closely agreed with MD and EPSR simulations. There is a pronounced decrease in the edge- and corner-shared In–Zn distances arising from the shorter Zn–O bond length, Zn–O tetrahedral coordination, and In–O–Zn angle in amorphous ZITO compared to its crystalline counterpart. A maximum in electrical mobility was observed for the amorphous film just before crystallization occurred. While the peak is broad, consistent with nearly unchanged overall cation–oxygen coordination in the amorphous films, ESPR results indicate that the tetrahedral coordination follows the conductivity trend.

Grazing Incidence X-ray↗