Search NASA⌕ Search

SEARCH · Search NASA

Results for “Enterprise Architecture”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

Integrating Cybersecurity Risk Assessment with Process Safety in Chemical Process Industries

This dissertation bridges the gap between industrial cybersecurity and traditional process safety by introducing an integrated Cyber-LOPA framework that combines the Purdue Enterprise Reference Architecture, Cyber Kill Chain, and CVSS v4.0 metrics. Demonstrated on a High-Density Polyethylene slurry process, the work illustrates how cyber threats targeting automation systems can bypass physical protection layers, proving the necessity of unified risk assessments to prevent cyber-induced physical incidents in chemical manufacturing plants.

Cyber-LOPA (CLOPA)↗

Toward a persistent event-streaming system for high-performance computing applications

High-performance computing (HPC) applications have traditionally relied on parallel file systems and file transfer services to manage data movement and storage. Alternative approaches have been proposed that use direct communications between application components, trading persistence and fault tolerance for speed. Event-driven architectures, as popularized in enterprise contexts, present a compelling middle ground, avoiding the performance cost and API constraints of parallel file systems while retaining persistence and offering impedance matching between application components. However, adapting streaming frameworks to HPC workloads requires addressing challenges unique to HPC systems. This paper investigates the potential for a streaming framework designed for HPC infrastructures and use cases. We introduce Mofka, a persistent event-streaming framework designed specifically for HPC environments. Mofka combines the capabilities of a traditional streaming service with optimizations tailored to the HPC context, such as support for massively multicore nodes, efficient scaling for large producer-consumer workflows, RDMA-enabled high-performance network communications, specialized network fabrics with multiple links per node, and efficient handling of large scientific data payloads. Built using the Mochi suite of HPC data service components, Mofka provides a lightweight, modular, and high-performance solution for persistent streaming in HPC systems. We present the architecture of Mofka and evaluate its performance against Kafka and Redpanda using benchmarks on diverse platforms, including Argonne's Polaris and Oak Ridge's Frontier supercomputers, showing up to 8× improvement in throughput in some scenarios. We then demonstrate its utility in several real-world applications: a tomographic reconstruction pipeline, a workflow for the discovery of metal-organic frameworks for carbon capture, and the instrumentation of Dask workflows for provenance tracking and performance analysis.

HPC↗

Designing a Comprehensive IDS Strategy for a Zero Trust Architecture Environment

Zero Trust Architecture or ZTA is a cybersecurity model for enterprises to structure their networked resources around to maintain total security externally and internally. In a Zero Trust environment, no part of the network is considered "trustworthy" and thus should be scrutinized and monitored extensively as is done in traditional "Trust But Verify" schemes at the network's perimeter. In this way, Zero Trust Architecture is a superior model for securing access to networked resources at the enterprise level. Fermilab, in pursuit of a better security posture, has decided to embrace this model of architecture for its network. Attaining this goal requires tremendous infrastructural, policy, and procedural adjustments that will affect all the lab's personnel and resources.

D'Antonio, Lucas↗

Future foundries: A convergent manufacturing platform

This article introduces the Future Foundries platform developed at Oak Ridge National Laboratory, a first-generation research system designed to demonstrate convergent manufacturing. Convergent manufacturing brings together additive, subtractive, and transformative processes in a digitally interconnected environment to enable end-to-end production workflows. By linking traditionally discrete steps, convergent platforms accelerate production, improve repeatability, and support high-mix, low-volume manufacturing. The Future Foundries platform exemplifies this vision in practice by combining four modular, vendor-agnostic process cells that include robotic WAAM, induction heating, optical metrology, and machining, coordinated through an automated pallet handler and a ROS 2-based digital thread. This architecture provides the flexibility and scalability needed for agile production in small and medium-sized manufacturing enterprises and for field deployable manufacturing. Two case studies illustrate the platform’s capabilities. The first presents an integrated workflow for fabricating, transforming, and repairing critical replacement components, showing how consolidated thermal, additive, inspection, and machining operations reduce manual part handling and streamline process flow. The second case study highlights coordinated multi-part production enabled by automated pallet logistics and multi-cell scheduling. Together, these examples showcase convergent manufacturing as a practical and scalable strategy for strengthening domestic casting and forging capacity, improving supply-chain resilience, and enabling rapid, adaptable production of mission-critical components.

Convergent manufacturing↗

Architectural Approaches for Integrating ADMS and DERMS: Challenges, Comparisons, and Real-World Use Cases

The electrical distribution landscape is rapidly transforming due to the proliferation of distributed energy resources (DERs) such as solar panels, wind turbines, battery storage systems, combined heat and power units, and electric vehicles, introducing variability and uncontrollability that traditional grid operators are ill-equipped to manage. This transformation is further accelerated by advancements in Information and Communication Technology infrastructure that connects control centers with end devices, demanding automation and a deeper understanding of new technologies by utility personnel. Advanced grid control techniques using system-level optimization, Artificial Intelligence, and Machine Learning at the enterprise level and distributed level are evolving to address these issues. There is also an opportunity to utilize the enormous data created by these new DER technologies in the grid. Advanced Distribution Management Systems (ADMS) and Distributed Energy Resource Management Systems (DERMS) are critical in addressing these challenges by automating grid operations and enhancing reliability. Given the relatively recent development of ADMS and DERMS, and the still relatively low level of ADMS and DERMS deployment in the industry, there is a notable deficiency in the comprehensive understanding of the challenges and benefits associated with these new technologies, especially with their complementary natures and integration architectures. This paper aims to bridge the knowledge gap in ADMS and DERMS integration, presenting three distinct integration architectures currently available, and discussing the challenges and benefits of each architecture to guide utilities, industry professionals, and researchers in optimizing grid management and decision-making processes for a resilient and efficient energy future.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Future Electric Power Industry and Grids—Now What Again is Our Destination?

Having a vision that others agree to support and work toward is highly desirable but hard to achieve. We seem to lack a common and shared understanding of the vision—or worse, multiple visions (vivid mental images or documented statements) with varying areas of focus and details: • some appear to be similar but have differing underlying goals and characteristics, or • some reflect differing viewpoints as to effects on various stakeholders. These desirable and undesirable situations apply to realizing visions for enterprise and industry, including the electricity sector. Multiple industry stakeholder groups have developed goals, industry vision statements, and characterizations of the future. The viewpoints are promoted, discussed, refined by their stakeholder group, and often published to promote broad understanding and to inform or influence others. The GridWise Architecture Council asked itself how well-aligned these characterizations of the future are. If these publications collectively set the overall direction for the industry, it is useful to identify their answers to questions such as, where is the electric industry headed, guided by what objectives, and with what role(s) for the customers, electric utilities, and other stakeholders? Are these goals, visions, and future states moving toward a common vision, do they provide value for the stakeholders, and are they likely to meet the objective stated? This paper addresses these questions via an assessment and characterization of a sampling of stakeholder groups’ publicly available vision and future state reports for the electricity industry. Identified electric power grid architectural topic areas needing further work are described, along with GridWise Architecture Council analysis and observation, potential collaborative work efforts, and suggested next steps.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Generative AI in Supply Chain Management: Applications, Challenges, and Future Directions

Supply chain management (SCM) is undergoing rapid transformation due to increasing global complexity, demand volatility, and operational disruptions. Generative Artificial Intelligence (GenAI) has emerged as a powerful paradigm capable of synthesizing data, simulating operational scenarios, and enabling adaptive decision-making across supply chain networks. This paper presents a survey of GenAI’s role in SCM, focusing on its applications in predictive analytics, autonomous logistics, and fraud detection. Unlike traditional AI systems that rely primarily on predictive analytics, GenAI models, including large language models, generative adversarial networks, and diffusion-based architectures, enable the creation of synthetic supply chain scenarios and autonomous optimization strategies. This survey provides (1) a taxonomy of GenAI techniques for supply chain applications, (2) a comparative analysis of generative AI approaches with traditional machine learning, reinforcement learning, and blockchain-based methods, and (3) a discussion of key challenges such as data privacy, interpretability, and integration with legacy enterprise systems. Furthermore, we outline open research problems and propose directions for future research toward autonomous, resilient, and sustainable AI-driven supply chains.

15 - GEOTHERMAL ENERGY↗

LC-Opt: Benchmarking Reinforcement Learning and Agentic AI for End-to-End Liquid Cooling Optimization in Data Centers

Liquid cooling is critical for thermal management in high-density data centers with the rising AI workloads. However, machine learning-based controllers are essential to unlock greater energy efficiency and reliability, promoting sustainability. We present LC-Opt, a Sustainable Liquid Cooling (LC) benchmark environment, for reinforcement learning (RL) control strategies in energy-efficient liquid cooling of high-performance computing (HPC) systems. Built on the baseline of a high-fidelity digital twin of Oak Ridge National Lab's Frontier Supercomputer cooling system, LC-Opt provides detailed Modelica-based end-to-end models spanning site-level cooling towers to data center cabinets and server blade groups. RL agents optimize critical thermal controls like liquid supply temperature, flow rate, and granular valve actuation at the IT cabinet level, as well as cooling tower (CT) setpoints through a Gymnasium interface, with dynamic changes in workloads. This environment creates a multi-objective real-time optimization challenge balancing local thermal regulation and global energy efficiency, and also supports additional components like a heat recovery unit (HRU). We benchmark centralized and decentralized multi-agent RL approaches, demonstrate policy distillation into decision and regression trees for interpretable control, and explore LLM-based methods that explain control actions in natural language through an agentic mesh architecture designed to foster user trust and simplify system management. LC-Opt democratizes access to detailed, customizable liquid cooling models, enabling the ML community, operators, and vendors to develop sustainable data center liquid cooling control solutions.

Naug, Avisek [Hewlett Packard Enterprise]↗

System-Level Integration of Modular Language Models for Real-Time Risk Assessment in Third-Party Risk Management Systems

Large enterprises typically rely on dedicated teams to govern and implement security measures throughout their supply chains, ensuring compliance with enterprise security procedures. There is a significant reliance on Third-Party Risk Management (TPRM) platforms, which often require complete, highly structured information from potential vendors. The review and compliance assurance processes are time- and labor intensive, often requiring several rounds of review between the supply chain security risk management teams, business users, and potential vendors, leading to delays in the supply chain processing and consumer experience. Significant challenges in the risk management paradigm include handling unstructured data in various formats and providing real-time feedback to users to reduce the required review time. This paper presents a novel solution to these challenges. A modular multi-step system architecture is proposed using advances in language processing, specifically for unstructured responses and provides real-time feedback (i.e., 3 seconds) so that users can improve their responses before the TPSRM team review. This novel system architecture will increase information accuracy and significantly reduce time and labor during the review process.

99 - GENERAL AND MISCELLANEOUS↗

Facies Analysis of the Prairie Du Chien Group in the Illinois Basin and Analogous Rocks in Missouri and Kentucky

Funded in 2023 by the U.S. Department of Energy’s Phase II Carbon Storage Assurance Facility Enterprise (CarbonSAFE) initiative, a Heidelberg Materials cement plant in Mitchell, Indiana, is currently being evaluated as a potential Carbon Capture and Storage (CCS) subsurface injection site. The Heidelberg CCS project targets the middle to upper Prairie du Chien Group (Early Ordovician) in southwestern Indiana. Assessment of reservoir feasibility requires collection of field data, seismic surveys, well-log correlation, geologic modeling, characterization well drilling, well testing, and reservoir simulation. However, the proposed Heidelberg CCS site is in a data-limited region, lacking both outcrop analogs and deep wells penetrating the target interval, which makes geologic modelling difficult prior to drilling a characterization well. To directly address this problem, the present study was undertaken to understand the sedimentologic composition and stratigraphic architecture of the Prairie du Chien Group from analogous outcrops and cores in the Illinois Basin and adjacent regions.

Ali, Shah Bilawal [Univ. of Illinois at Urbana-Cha↗

Lustre Unveiled: Evolution, Design, Advancements, and Current Trends

The Lustre filesystem serves as a vital element in high-performance parallel storage, meeting the rising demands of scientific, research, and enterprise environments. Widely deployed across HPC environments, ranging from small-scale applications in AI/ML, to domains like oil and gas, drug discovery, and meteorology, and manufacturing, Lustre addresses the universal challenge of efficiently accessing vast and ever-increasing volumes of data. Lustre is the filesystem of choice on six out of the top 10 fastest supercomputers in the world today, over 65% of the top 100, and also for over 60% of the top 500. Despite its widespread popularity, there is a lack of a complete and up-to-date reference, covering Lustre’s evolution, design, and various advancements made over the years. In this journal, we aim to fill this gap by providing a comprehensive journey of Lustre, including its history with significant contributions to HPC, detailed architecture and design elements, exploration of advancements added through its evolution, and future directions. Additionally, we present a comparison of Lustre with other prominent storage technologies of the era. To illustrate the current state of Lustre, we analyze several filesystem trends, including utilization, performance, and usage patterns on Orion, the Lustre filesystem on the first exascale supercomputer Frontier. We hope that this journal serves as a comprehensive educational reference for the current and future generations interested in HPC filesystem storage aspects.

97 MATHEMATICS AND COMPUTING↗

Preparing MPICH for exascale

The advent of exascale supercomputers heralds a new era of scientific discovery, yet it introduces significant architectural challenges that must be overcome for MPI applications to fully exploit its potential. Among these challenges is the adoption of heterogeneous architectures, particularly the integration of GPUs to accelerate computation. Additionally, the complexity of multithreaded programming models has also become a critical factor in achieving performance at scale. The efficient utilization of hardware acceleration for communication, provided by modern NICs, is also essential for achieving low latency and high throughput communication in such complex systems. In response to these challenges, the MPICH library, a high-performance and widely used Message Passing Interface (MPI) implementation, has undergone significant enhancements. Here, this paper presents four major contributions that prepare MPICH for the exascale transition. First, we describe a lightweight communication stack that leverages the advanced features of modern NICs to maximize hardware acceleration. Second, our work showcases a highly scalable multithreaded communication model that addresses the complexities of concurrent environments. Third, we introduce GPU-aware communication capabilities that optimize data movement in GPU-integrated systems. Finally, we present a new datatype engine aimed at accelerating the use of MPI derived datatypes on GPUs. These improvements in the MPICH library not only address the immediate needs of exascale computing architectures but also set a foundation for exploiting future innovations in high-performance computing. By embracing these new designs and approaches, MPICH-derived libraries from HPE Cray and Intel were able to achieve real exascale performance on OLCF Frontier and ALCF Aurora respectively.

Guo, Yanfei [Argonne National Laboratory (ANL), Ar↗