Search NASA⌕ Search

SEARCH · Search NASA

Results for “data access”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 415 records · Page 23

A Simple XML Producer-Consumer Protocol

There are many different projects from government, academia, and industry that provide services for delivering events in distributed environments. The problem with these event services is that they are not general enough to support all uses and they speak different protocols so that they cannot interoperate. We require such interoperability when we, for example, wish to analyze the performance of an application in a distributed environment. Such an analysis might require performance information from the application, computer systems, networks, and scientific instruments. In this work we propose and evaluate a standard XML-based protocol for the transmission of events in distributed systems. One recent trend in government and academic research is the development and deployment of computational grids. Computational grids are large-scale distributed systems that typically consist of high-performance compute, storage, and networking resources. Examples of such computational grids are the DOE Science Grid, the NASA Information Power Grid (IPG), and the NSF Partnerships for Advanced Computing Infrastructure (PACIs). The major effort to deploy these grids is in the area of developing the software services to allow users to execute applications on these large and diverse sets of resources. These services include security, execution of remote applications, managing remote data, access to information about resources and services, and so on. There are several toolkits for providing these services such as Globus, Legion, and Condor. As part of these efforts to develop computational grids, the Global Grid Forum is working to standardize the protocols and APIs used by various grid services. This standardization will allow interoperability between the client and server software of the toolkits that are providing the grid services. The goal of the Performance Working Group of the Grid Forum is to standardize protocols and representations related to the storage and distribution of performance data. These standard protocols and representations must support tasks such as profiling parallel applications, monitoring the status of computers and networks, and monitoring the performance of services provided by a computational grid. This paper describes a proposed protocol and data representation for the exchange of events in a distributed system. The protocol exchanges messages formatted in XML and it can be layered atop any low-level communication protocol such as TCP or UDP Further, we describe Java and C++ implementations of this protocol and discuss their performance. The next section will provide some further background information. Section 3 describes the main communication patterns of our protocol. Section 4 describes how we represent events and related information using XML. Section 5 describes our protocol and Section 6 discusses the performance of two implementations of the protocol. Finally, an appendix provides the XML Schema definition of our protocol and event information.

Smith, Warren↗

Crew Health and Performance Integrated Data Architecture Project Update

Future Human Exploration missions will face new constraints as crews move further from terrestrial communication, resupply, and the real-time support enjoyed by Low Earth Orbit missions today. Exploration crews will need to be more self-reliant and able to respond to emergencies without immediate support from ground-based personnel. A new generation of technologies, employing advanced analytical and predictive modeling techniques, is needed to assist the crew’s work, help maintain their health, and inform the decisions they make on these future Exploration missions. The Crew Health and Performance Integrated Data Architecture (CHP-IDA) project is laying a foundation for these future technologies by integrating sources of data generated by and around the crew then providing them though common data models and Application Programming Interfaces to external systems. The combined data model makes comprehensive Crew Health and Performance data accessible and more meaningful to the decision-making process. This presentation will describe the currently ongoing effort to develop a path-to-flight concept of the CHP-IDA software, current integrations that have been developed, updates to the architecture, and examples of the corresponding exploration scenarios in which CHP-IDA would be used.

Exploration↗

Hierarchical and Parallelizable Direct Volume Rendering for Irregular and Multiple Grids

A general volume rendering technique is described that efficiently produces images of excellent quality from data defined over irregular grids having a wide variety of formats. Rendering is done in software, eliminating the need for special graphics hardware, as well as any artifacts associated with graphics hardware. Images of volumes with about one million cells can be produced in one to several minutes on a workstation with a 150 MHz processor. A significant advantage of this method for applications such as computational fluid dynamics is that it can process multiple intersecting grids. Such grids present problems for most current volume rendering techniques. Also, the wide range of cell sizes (by a factor of 10,000 or more), which is typical of such applications, does not present difficulties, as it does for many techniques. A spatial hierarchical organization makes it possible to access data from a restricted region efficiently. The tree has greater depth in regions of greater detail, determined by the number of cells in the region. It also makes it possible to render useful 'preview' images very quickly (about one second for one-million-cell grids) by displaying each region associated with a tree node as one cell. Previews show enough detail to navigate effectively in very large data sets. The algorithmic techniques include use of a kappa-d tree, with prefix-order partitioning of triangles, to reduce the number of primitives that must be processed for one rendering, coarse-grain parallelism for a shared-memory MIMD architecture, a new perspective transformation that achieves greater numerical accuracy, and a scanline algorithm with depth sorting and a new clipping technique.

Wilhelms, Jane↗

Internal NASA Study: NASAs Protoflight Research Initiative

The NASA Protoflight Research Initiative is an internal NASA study conducted within the Office of the Chief Engineer to better understand the use of Protoflight within NASA. Extensive literature reviews and interviews with key NASA members with experience in both robotic and human spaceflight missions has resulted in three main conclusions and two observations. The first conclusion is that NASA's Protoflight method is not considered to be "prescriptive." The current policies and guidance allows each Program/Project to tailor the Protoflight approach to better meet their needs, goals and objectives. Second, Risk Management plays a key role in implementation of the Protoflight approach. Any deviations from full qualification will be based on the level of acceptable risk with guidance found in NPR 8705.4. Finally, over the past decade (2004 - 2014) only 6% of NASA's Protoflight missions and 6% of NASA's Full qualification missions experienced a publicly disclosed mission failure. In other words, the data indicates that the Protoflight approach, in and of it itself, does not increase the mission risk of in-flight failure. The first observation is that it would be beneficial to document the decision making process on the implementation and use of Protoflight. The second observation is that If a Project/Program chooses to use the Protoflight approach with relevant heritage, it is extremely important that the Program/Project Manager ensures that the current project's requirements falls within the heritage design, component, instrument and/or subsystem's requirements for both the planned and operational use, and that the documentation of the relevant heritage is comprehensive, sufficient and the decision well documented. To further benefit/inform this study, a recommendation to perform a deep dive into 30 missions with accessible data on their testing/verification methodology and decision process to research the differences between Protoflight and Full Qualification missions' Design Requirements and Verification & Validation (V&V) (without any impact or special request directly to the project).

Protoflight↗

Policies and Procedures for Accessing Archived NASA Data via the Web

The National Space Science Data Center (NSSDC) was established by NASA to provide for the preservation and dissemination of scientific data from NASA missions. This white paper will address the NSSDC policies that govern data preservation and dissemination and the various methods of accessing NSSDC-archived data via the web.

James, Nathan↗

Preservation of Provenance and Context to Ensure Future Understandability of Airborne Earth Observations and Derived Data Products

Open-source science goes beyond making data from scientific projects (e.g., on-orbit/satellite missions, airborne and field investigations, and other data producing activities) openly available after they are generated, but involves and open sharing of information throughout the project lifecycle. Preservation of the data and associated information required for understanding and reusing the data well after the scientific projects is a contributor to open-source science as well. Considering the high investment in the on-orbit/satellite missions, we had developed a document titled “NASA Earth Science Data Preservation Content Specification (PCS)” in 2011. This document has been used as a requirement for recent on-orbit/satellite missions by NASA. Recently it became clear that the specifications should be applied to other scientific projects as well. Therefore, the document was revised to cover other types of projects, and a Preservation Content Implementation Guidance (PCIG) document was also developed. The revised PCS, and the PCIG, were published in 2022. The purpose of this presentation is to highlight the contents of these documents as they apply to suborbital/airborne investigations. The PCS calls for content preservation in eight general categories - Measuring Instrument/Platform Description, Instrument and Science Data Products and Metadata, Science Raw Data, Product and Algorithm Documentation, Instrument Calibration, Science Algorithm Software, Science Data Product Algorithm Inputs, Science Data Product Validation, and Science Data Access and Analysis Tools. While all these categories apply to various types of projects, a few clarifying sentences have been added to the descriptions of contents in each of the categories to show which categories are especially important to airborne and field investigations and where some contents are not applicable (or difficult to obtain). The PCIG document provides some general guidance applicable to all types of projects and specific guidance in a separate section for airborne and field investigations. This section calls out typical artifacts produced during such investigations that can meet the spirit of the various PCS categories.

remote sensing↗

A strainmeter array as the fulcrum of novel observatory sites along the Alto Tiberina Near Fault Observatory

Fault slip is a complex natural phenomenon involving multiple spatiotemporal scales from seconds to days to weeks. To understand the physical and chemical processes responsible for the full fault slip spectrum, a multidisciplinary approach is highly recommended. The Near Fault Observatories (NFOs) aim at providing high-precision and spatiotemporally dense multidisciplinary near-fault data, enabling the generation of new original observations and innovative scientific products. The Alto Tiberina Near Fault Observatory is a permanent monitoring infrastructure established around the Alto Tiberina fault (ATF), a 60 km long low-angle normal fault (mean dip 20°), located along a sector of the Northern Apennines (central Italy) undergoing an extension at a rate of about 3 mm yr –1 . The presence of repeating earthquakes on the ATF and a steep gradient in crustal velocities measured across the ATF by GNSS stations suggest large and deep (5–12 km) portions of the ATF undergoing aseismic creep. Both laboratory and theoretical studies indicate that any given patch of a fault can creep, nucleate slow earthquakes, and host large earthquakes, as also documented in nature for certain ruptures (e.g., Iquique in 2014, Tōhoku in 2011, and Parkfield in 2004). Nonetheless, how a fault patch switches from one mode of slip to another, as well as the interaction between creep, slow slip, and regular earthquakes, is still poorly documented by near-field observation. With the strainmeter array along the Alto Tiberina fault system (STAR) project, we build a series of six geophysical observatory sites consisting of 80–160 m deep vertical boreholes instrumented with strainmeters and seismometers as well as meteorological and GNSS antennas and additional seismometers at the surface. By covering the portions of the ATF that exhibits repeated earthquakes at shallow depth (above 4 km) with these new observatory sites, we aim to collect unique open-access data to answer fundamental questions about the relationship between creep, slow slip, dynamic earthquake rupture, and tectonic faulting.

58 GEOSCIENCES↗

NPSS on NASA's IPG: Using CORBA and Globus to Coordinate Multidisciplinary Aeroscience Applications

Within NASA's High Performance Computing and Communication (HPCC) program, the NASA Glenn Research Center is developing an environment for the analysis/design of aircraft engines called the Numerical Propulsion System Simulation (NPSS). The vision for NPSS is to create a "numerical test cell" enabling full engine simulations overnight on cost-effective computing platforms. To this end, NPSS integrates multiple disciplines such as aerodynamics, structures, and heat transfer and supports "numerical zooming" between O-dimensional to 1-, 2-, and 3-dimensional component engine codes. In order to facilitate the timely and cost-effective capture of complex physical processes, NPSS uses object-oriented technologies such as C++ objects to encapsulate individual engine components and CORBA ORBs for object communication and deployment across heterogeneous computing platforms. Recently, the HPCC program has initiated a concept called the Information Power Grid (IPG), a virtual computing environment that integrates computers and other resources at different sites. IPG implements a range of Grid services such as resource discovery, scheduling, security, instrumentation, and data access, many of which are provided by the Globus toolkit. IPG facilities have the potential to benefit NPSS considerably. For example, NPSS should in principle be able to use Grid services to discover dynamically and then co-schedule the resources required for a particular engine simulation, rather than relying on manual placement of ORBs as at present. Grid services can also be used to initiate simulation components on parallel computers (MPPs) and to address inter-site security issues that currently hinder the coupling of components across multiple sites. These considerations led NASA Glenn and Globus project personnel to formulate a collaborative project designed to evaluate whether and how benefits such as those just listed can be achieved in practice. This project involves firstly development of the basic techniques required to achieve co-existence of commodity object technologies and Grid technologies; and secondly the evaluation of these techniques in the context of NPSS-oriented challenge problems. The work on basic techniques seeks to understand how "commodity" technologies (CORBA, DCOM, Excel, etc.) can be used in concert with specialized "Grid" technologies (for security, MPP scheduling, etc.). In principle, this coordinated use should be straightforward because of the Globus and IPG philosophy of providing low-level Grid mechanisms that can be used to implement a wide variety of application-level programming models. (Globus technologies have previously been used to implement Grid-enabled message-passing libraries, collaborative environments, and parameter study tools, among others.) Results obtained to date are encouraging: we have successfully demonstrated a CORBA to Globus resource manager gateway that allows the use of CORBA RPCs to control submission and execution of programs on workstations and MPPs; a gateway from the CORBA Trader service to the Grid information service; and a preliminary integration of CORBA and Grid security mechanisms. The two challenge problems that we consider are the following: 1) Desktop-controlled parameter study. Here, an Excel spreadsheet is used to define and control a CFD parameter study, via a CORBA interface to a high throughput broker that runs individual cases on different IPG resources. 2) Aviation safety. Here, about 100 near real time jobs running NPSS need to be submitted, run and data returned in near real time. Evaluation will address such issues as time to port, execution time, potential scalability of simulation, and reliability of resources. The full paper will present the following information: 1. A detailed analysis of the requirements that NPSS applications place on IPG. 2. A description of the techniques used to meet these requirements via the coordinated use of CORBA and Globus. 3. A description of results obtained to date in the first two challenge problems.

Lopez, Isaac↗

Towards the Interoperability of Web, Database, and Mass Storage Technologies for Petabyte Archives

At the San Diego Supercomputer Center, a massive data analysis system (MDAS) is being developed to support data-intensive applications that manipulate terabyte sized data sets. The objective is to support scientific application access to data whether it is located at a Web site, stored as an object in a database, and/or storage in an archival storage system. We are developing a suite of demonstration programs which illustrate how Web, database (DBMS), and archival storage (mass storage) technologies can be integrated. An application presentation interface is being designed that integrates data access to all of these sources. We have developed a data movement interface between the Illustra object-relational database and the NSL UniTree archival storage system running in a production mode at the San Diego Supercomputer Center. With this interface, an Illustra client can transparently access data on UniTree under the control of the Illustr DBMS server. The current implementation is based on the creation of a new DBMS storage manager class, and a set of library functions that allow the manipulation and migration of data stored as Illustra 'large objects'. We have extended this interface to allow a Web client application to control data movement between its local disk, the Web server, the DBMS Illustra server, and the UniTree mass storage environment. This paper describes some of the current approaches successfully integrating these technologies. This framework is measured against a representative sample of environmental data extracted from the San Diego Ba Environmental Data Repository. Practical lessons are drawn and critical research areas are highlighted.

Moore, Reagan↗

The Evolution of Randomized Clinical Trial Designs to Assess Therapeutics in Alzheimer Disease

Importance The success of recent randomized clinical trials (RCTs) for Alzheimer disease (AD), particularly those focusing on anti-amyloid therapies, has been discussed at length. However, the evolution of RCT design features for AD that preceded this success remain underexplored. Objective To describe temporal changes in the features of RCT design for interventions in AD. Evidence Review PubMed, Scopus, and Web of Science databases were searched in January 2025 for phase 2 and 3 AD RCTs published between January 1992 and December 2024. RCTs that investigated an intervention for AD, with a placebo or standard-of-care control group, were included. Four assessors independently reviewed full-text articles to capture study characteristics. Main Outcomes and Measures The number of participants and the duration of RCTs as well as the target population, outcomes, and funding were extracted from published reports. These features were analyzed with respect to time using linear regression and χ 2 analyses. Results The study included 203 RCTs with 79 589 participants testing interventions in AD. From 1992 to 2024, the mean sample size increased by 464% for phase 2 RCTs (from 42 to 237), and 50% for phase 3 RCTs (from 632 to 951), while the mean trial duration increased by 188% (from 16 to 46 weeks) for phase 2, and 256% (from 20 to 71 weeks) for phase 3 RCTs. This longer duration of RCTs may be partially attributed by a greater share of disease-modifying rather than symptomatic treatments. Similarly, more recent trials required AD biomarker evidence for enrollment (from 1 of 36 [2.7%] before 2006 to 40 of 76 [52.6%] since 2019). A substantial difference in the type of therapeutics researched was observed, with anti-amyloid and anti-tau RCTs being more likely to be funded by the pharmaceutical industry compared with neurotransmitter or other RCTs (anti-amyloid or anti-tau, 68 of 71 [95.8%]; neurotransmitter, 52 of 69 [77.6%]; other, 33 of 52 [63.5%]). RCT transparency improved, with more frequent data accessibility statements, registered reports, and better reporting on race and ethnicity. Conclusions and Relevance This methodology research of AD RCTs highlights substantial changes in key features of AD clinical trials from 1992 to 2024. AD RCTs have become larger and longer, such that they are powered to detect smaller clinical differences. The increased sample sizes and duration should enable the detection of smaller and more slowly occurring outcomes, which may lead to successful RCTs of therapies with slower and more subtle efficacy.

General & Internal Medicine↗

Analysis Facilities for the HL-LHC White Paper

This white paper presents the current status of the R&D for Analysis Facilities (AFs) and attempts to summarize the views on the future direction of these facilities. These views have been collected through the High Energy Physics (HEP) Software Foundation’s (HSF) Analysis Facilities forum (HSF Analysis Facilities Forum), established in March 2022, the Analysis Ecosystems II workshop (Analysis Ecosystems Workshop II), that took place in May 2022, and the WLCG/HSF pre-CHEP workshop (WLCG–HSF pre-CHEP Workshop), that took place in May 2023. The paper attempts to cover all the aspects of an analysis facility.

97 MATHEMATICS AND COMPUTING↗

A Comprehensive Chemistry Evaluation and Diagnostics Package for E3SM – ChemDyg Version 1.1.0

The Chemistry Evaluation and Diagnostics Package (ChemDyg) is an open-source tool designed for the Energy Exascale Earth System Model (E3SM) developed by the U.S. Department of Energy. ChemDyg facilitates routine evaluation, tailored development, and in-depth analysis of atmospheric chemistry through its modular architecture, allowing users to compare model outputs with observational data. Version 1.1.0 introduces a robust set of diagnostic capabilities, including climatology, time evolution of key tracers, diurnal and annual cycle analyses, and extensive budget diagnostics. These features help identify model discrepancies and enhance the representation of atmospheric chemistry in E3SM. Each self-contained diagnostic set includes dedicated scripts and documentation for ease of use. The interactive HTML output improves data accessibility, accelerating chemistry model development. Additionally, ChemDyg's flexible framework allows for customization, enabling users to create unique diagnostic sets for specific scientific contributions.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

DTLMod: A simulation framework for in situ workflow optimization

In situ processing workflows have become essential for coping with the explosion in data volume and velocity in large-scale scientific computing, providing domain scientists with early insights at runtime. Multiple frameworks implement this paradigm through a data transport layer (DTL), offering different data access modes and deployment schemes, but researchers currently lack the appropriate tools to assess design and deployment options before committing to costly real experiments. We introduce DTLMod, an open-source simulated DTL that enables performance evaluation of in situ workflow configurations at scale. Built on SimGrid, it links into any SimGrid-based simulator and is available in C++ and Python. We evaluate DTLMod along four axes: scalability (tens of thousands of simulated processes across interconnected clusters in seconds, with linear memory scaling), versatility (three implementation variants trading fidelity for speed), accuracy (simulated times faithfully reflecting real behavior), and practical utility (two use cases demonstrating evidence-based workflow design decisions).

Suter, Fred [ORNL] (ORCID:0000000319021955)↗

Barriers to adopting artificial intelligence and machine learning technologies in nuclear power

Artificial intelligence and machine learning (AI/ML) technologies offer unique opportunities to transform nuclear plant operations and power generation. Benefits will be felt not only within existing analog and digital instrumentation and control, but also within work processes, the integration of people with technology and most importantly, the business case. The application of this new technology can help simplify complex problems and produce more effective decision-making, making nuclear power safer, more efficient, and more economically viable in the current energy market. Nonetheless, there are potential barriers to its adoption that must be overcome. The purpose of this paper is to categorize, review, and discuss barriers to AI/ML adoption within the nuclear power industry, with a focus on existing commercial reactors. Unique considerations for advanced reactors are also offered. Here we provide a comprehensive overview of the historical, technical, and business barriers that the industry faces, as well as stakeholder readiness, and end-user acceptance. We underscore the importance of user experience and offer potential solutions in overcoming each barrier. These include provisions for easier plant data access, a friendly regulatory environment, and investment in user trust and explainable AI.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

Toward Unified Autonomous Scattering Experiments: A Cross-Facility Case Study at ALS and PETRA III

Autonomous experiments rely on the integration of control, data acquisition, analysis, and decision-making frameworks. While such systems have been demonstrated at individual facilities, adapting them to additional instruments remains challenging due to differences in local infrastructure. We present a modular workflow that connects existing open-source tools for data access (Tiled), workflow orchestration (Prefect), analysis and visualization (pyFAI, Plotly Dash), and Gaussian-process-based adaptive sampling (gpCAM) into a unified framework for autonomous scattering experiments. The same configuration operates across two synchrotron beamlines (ALS 7.3.3 and PETRA III P03) with only minimal facility-specific adjustments, as shown in proof-of-concept demonstrations. This validates that a consistent design emphasizing modularity and shared interfaces can ease deployment across diverse experimental environments. The resulting framework provides a flexible foundation for extending autonomous control and analysis capabilities beyond a single beamline or instrument.

47 OTHER INSTRUMENTATION↗

Scalable fabrication of an array-type fixed-target device for automated room temperature X-ray protein crystallography

X-ray crystallography is one of the leading tools to analyze the 3-D structure, and therefore, function of proteins and other biological macromolecules. Traditional methods of mounting individual crystals for X-ray diffraction analysis can be tedious and result in damage to fragile protein crystals. Furthermore, the advent of multi-crystal and serial crystallography methods explicitly require the mounting of larger numbers of crystals. To address this need, we have developed a device that facilitates the straightforward mounting of protein crystals for diffraction analysis, and that can be easily manufactured at scale. Inspired by grid-style devices that have been reported in the literature, we have developed an X-ray compatible microfluidic device that can be used to trap protein crystals in an array configuration, while also providing excellent optical transparency, a low X-ray background, and compatibility with the robotic sample handling and environmental controls used at synchrotron macromolecular crystallography beamlines. At the Stanford Synchrotron Radiation Lightsource (SSRL), these capabilities allow for fully remote-access data collection at controlled humidity conditions. Furthermore, we have demonstrated continuous manufacturing of these devices via roll-to-roll fabrication to enable cost-effective and efficient large-scale production.

chemical engineering↗

Migration to WebDAV in Belle II Experiment

The usage of WebDAV protocol has become more and more popular within the physics experiments using grid middleware in the last decade, and today it represents a valid alternative to the GridFTP currently supported at best-effort level after the retirement of Globus Toolkit. Belle II experiment established the adoption of WebDAV protocol as the main protocol for data access and third-party-copy transfers, without relying on Storage Resource Manager interface (SRM). The migration process, carried on with continuous and gradual steps, has required a large effort to guarantee a smooth transition maintaining the production infrastructure fully operational. In this contribution we show the transition process, the tool of support developed to monitor step by step the status of third-party-copy support with WebDAV protocol by storages of the collaboration tested in both case pull and push, the strategy adopted to configure DIRAC and the solutions put in place for the corner cases. Finally, we will present some statistics of utilization and we will analyse the achieved results.

97 MATHEMATICS AND COMPUTING↗

Modeling Distributed Computing Infrastructures for HEP Applications

Predicting the performance of various infrastructure design options in complex federated infrastructures with computing sites distributed over a wide area network that support a plethora of users and workflows, such as the Worldwide LHC Computing Grid (WLCG), is not trivial. Due to the complexity and size of these infrastructures, it is not feasible to deploy experimental test-beds at large scales merely for the purpose of comparing and evaluating alternate designs. An alternative is to study the behaviours of these systems using simulation. This approach has been used successfully in the past to identify efficient and practical infrastructure designs for High Energy Physics (HEP). A prominent example is the Monarc simulation framework, which was used to study the initial structure of the WLCG. New simulation capabilities are needed to simulate large-scale heterogeneous computing systems with complex networks, data access and caching patterns. A modern tool to simulate HEP workloads that execute on distributed computing infrastructures based on the SimGrid and WRENCH simulation frameworks is outlined. Studies of its accuracy and scalability are presented using HEP as a case-study. Hypothetical adjustments to prevailing computing architectures in HEP are studied providing insights into the dynamics of a part of the WLCG and candidates for improvements.

Horzela, Maximilian↗