Search NASASearch

SEARCH · Search NASA

Results for “multiple agents”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Design, Preparation, and Execution of the 100-AV Field Test for the CIRCLES Consortium: Methodology and Implementation of the Largest Mobile Traffic Control Experiment to Date

This article presents the comprehensive design, setup, execution, and evaluation of the MegaVanderTest (MVT) experiment conducted by the Congestion Impacts Reduction via CAV-in-the-Loop Lagrangian Energy Smoothing (CIRCLES) Consortium, which aimed to mitigate traffic congestion using partially autonomous vehicles (AVs) (see “Summary”). The experiment involved 100 vehicles on Nashville’s Interstate 24 (I-24) highway, utilizing various control algorithms to smooth stop-and-go traffic waves. The execution of the MVT experiment required a coordinated effort from multiple teams. This article details the meticulous planning process, the coordinated efforts of multiple teams, and the innovative use of a dynamic agent-based simulation framework for traffic evaluation. Here, the contributions of this work include demonstrating and providing a detailed roadmap for large-scale live traffic experiments, illustrating the lessons learned from the MVT experiment, and introducing the other articles in this issue and their complementary relationship in the MVT experiment.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI

Adaptive Reinforcement Learning Control for Power Distribution in Multi-Output Resonant Converters

This paper presents an adaptive reinforcement learning (ARL)-based control framework for efficient power distribution in a multi-output resonant converter for UAV applications. The proposed system is based on a high-frequency isolated resonant architecture, where a single energy source supplies multiple propulsion loads through independently controlled output rectifiers, addressing the need for coordinated multi-motor power management. The ARL framework dynamically allocates output power by learning optimal phase-shift control actions under varying load demands and operating conditions. The agent autonomously determines control parameters that maximize conversion efficiency while ensuring accurate power sharing among multiple outputs. In addition, the proposed approach enables adaptive operation without requiring detailed system modeling or manual tuning. Experimental results demonstrate stable and efficient performance over a wide range of operating conditions, confirming the effectiveness and robustness of the learning-based control strategy for multi-output resonant converter system.

Asa, Erdem [ORNL] (ORCID:0000000190884812)

Optimal CO 2 storage management considering safety constraints in multi-stakeholder multi-site GCS projects: A Markov game perspective

Geological carbon storage (GCS) projects could involve a diverse array of stakeholders or players from public, private, and regulatory sectors, each with different objectives and responsibilities. Given the complexity, scale, and long-term nature of GCS operations, determining whether individual stakeholders can independently optimize their interests — or whether collaborative coalition agreements are needed — remains a central question for effective GCS project planning and management. To access large, high-quality storage resources, future GCS deployment may increasingly occur in geologically connected sites, where shared geological features such as pressure space and reservoir pore capacity can lead to competitive behavior among stakeholders. In this work, we propose a paradigm based on Markov games to quantitatively investigate how different coalition structures affect the goals of stakeholders. We frame this multi-stakeholder multi-site problem as a multi-agent reinforcement learning problem with safety constraints. Our approach enables agents to learn optimal strategies while complying with safety regulations. We present an example where multiple operators are injecting CO 2 into their respective project areas in a geologically connected basin. To address the high computational cost of repeated simulations of high fidelity models, a previously developed surrogate model based on the Embed-to-Control (E2C) framework is employed. Our results demonstrate the effectiveness of the proposed framework in addressing optimal management of CO 2 storage when multiple stakeholders with different objectives and goals are involved.

58 GEOSCIENCES

Reinforcement Learning Control for Buildings Co-Optimizing Energy, Comfort, and Indoor Air Quality: An Annual Assessment

Efficient control of Heating, Ventilation, and Air Conditioning (HVAC) systems is crucial for optimizing energy use and maintaining indoor comfort in buildings. Traditional control methods, such as PID control, cannot handle energy use trade-offs among multiple components in the building energy system at a supervisory level. Reinforcement learning (RL) presents a promising solution, offering adaptive and data-driven control strategies that optimize performance over time. However, RL also faces several challenges, including the conflicts encountered in co-optimizing energy savings, occupant comfort, and indoor air quality, and the requirement for extensive interactions with the environment in training. We proposed a flexible simulation platform that integrates a hybrid model for RL training and designed an RL agent to control the entire central HVAC system, focusing on co-optimizing energy consumption, thermal comfort, and indoor air quality ($\text{CO}_{2}$ and PM2.5 concentrations). Finally, we evaluated the RL agent's performance over an annual cycle. Our findings indicate that the RL agent can effectively manage the HVAC system with 14.7 % energy savings annually and balance multiple objectives, which demonstrates significant potential for improving HVAC system control and sustainability in buildings.

Guo, Fangzhou

Agent-based modeling for multimodal transportation of CO 2 for carbon capture, utilization, and storage: CCUS-agent

Here, to understand the system-level interactions between the entities in Carbon Capture, Utilization, and Storage (CCUS), an agent-based foundational modeling tool, CCUS-Agent, is developed for a large-scale study of transportation flows and infrastructure in the United States. Key features of the tool include (i) modular design, (ii) multiple transportation modes, (iii) capabilities for extension, and (iv) testing against various system components and networks of small and large sizes. Five matching algorithms for CO 2 supply agents (e.g., powerplants and industrial facilities) and demand agents (e.g., storage and utilization sites) are explored: Most Profitable First Year (MPFY), Most Profitable All Years (MPAY), Shortest Total Distance First Year (SDFY), Shortest Total Distance All Years (SDAY), and Shortest distance to long-haul transport All Years (ACAY). Before matching, the supply agent, demand agent, and route must be available, and the connection must be profitable. A profitable connection means the supply agent portion of revenue from the 45Q tax credit must cover the supply agent costs and all transportation costs, while the demand agent revenue portion must cover all demand agent costs. A case study employing over 5500 supply and demand agents and multimodal CCUS transportation infrastructure in the contiguous United States is conducted. The results suggest that it is possible to capture over 9 billion tonnes (GT) of CO 2 from 2025 to 2043, which will increase significantly to 22 GT if the capture costs are reduced by 40 %. The MPFY and SDFY algorithms capture more CO 2 earlier in the time horizon, while the MPAY and SDAY algorithms capture more later in the time horizon.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI

Zooming in on virtual commutes: Telecommuting impacts on mobility and sustainability

Motivated by a societal shift towards remote work and rapid advancement in information and communication technologies, this study examines the impact of telecommuting on urban mobility across the nine counties of the San Francisco Bay Area, California. Utilizing a behaviorally realistic integrated agent-based transportation model and recent data on telecommuting patterns, we simulate the impact of multiple telecommuting scenarios on transportation system outcomes. This provides a comprehensive picture of the impact of remote work on travel costs and accessibility of all travelers, factoring in changes in mode use and congestion. We analyze how telecommuting influences broader societal factors such as transportation energy consumption. Our findings indicate that increased telecommuting reduces overall person miles traveled and transportation energy consumption. Furthermore, telecommuting results in externality benefits by improving accessibility and reducing commute times for non-telecommuters.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI

ReLIC: Full-Scale Realization of Reinforcement Learning for Infrastructure Control

Prior efforts have shown that deep reinforcement learning (DRL) may provide a new method for controlling networked power systems. Though successful, prior approaches have not yet demonstrated their behavior on systems of realistic scale. This effort examined multiple theoretical and technical approaches to allow a DRL model to operate over a system of 2,000 buses or more. We find that allowing the DRL models to run training episodes in parallel provides near limitless efficiency gains, allowing us to train successful agents to behave on our Kuramoto transmission model of up to 4,000 buses. We further show that we can expand our PowerWorld DRL implementation to systems of up to 25 buses but struggle to go beyond this limit due to PowerWorld’s inability to run multiple instances at once. Finally, we examine a multi-agent approach and find that it performs as well if not better than our existing centralized approach.

97 MATHEMATICS AND COMPUTING

mada-tools: MCP servers, configurations, skills, and examples for MADA

MADA-tools (Multi-Agent Design Assistant tools) is a library for defining MCP (Model Context Protocol) servers that can be used by AI agents in the MADA project. Each MCP server provides a focused set of tools that enhances an LLM's knowledge and capabilities for a specific domain, for example, how to launch jobs with Flux versus Slurm. The library makes it easy to configure and start multiple MCP servers using configuration files or command line options. Once running, these servers are intended to be consumed by one or more agents in the MADA ecosystem. The system is designed to be extensible so that future projects can contribute their own MCP servers, skills, and toolsets.

Gunnarson, BrianS [Lawrence Livermore National Lab

Nuclear microreactor transient and load-following control with deep reinforcement learning

The economic feasibility of nuclear microreactors will depend on minimizing operating costs through advancements in autonomous control, especially when these microreactors are operating alongside other types of energy systems (e.g., renewable energy). This study explores the application of deep reinforcement learning (RL) for real-time drum control in microreactors, exploring performance in regard to load-following scenarios. By leveraging a point kinetics model with thermal and xenon feedback, we first establish a baseline using a single-output RL agent, then compare it against a traditional proportional–integral–derivative (PID) controller. This study demonstrates that RL controllers, including both single- and multi-agent RL (MARL) frameworks, can achieve similar or even superior load-following performance as traditional PID control across a range of load-following scenarios. In short transients, the RL agent was able to reduce the tracking error rate in comparison to PID by one half to one third. Over extended 300-minute load-following scenarios in which xenon feedback becomes a dominant factor, PID maintained better accuracy, but RL still remained within a 1% error margin despite being trained only on short-duration scenarios. This highlights RL’s strong ability to generalize and extrapolate to longer, more complex transients, affording substantial reductions in training costs and reduced overfitting. Furthermore, when control was extended to multiple drums, MARL enabled independent drum control as well as maintained reactor symmetry constraints without sacrificing performance---an objective that standard single-agent RL could not learn. We also found that, as increasing levels of Gaussian noise were added to the power measurements, the RL controllers were able to maintain lower error rates than PID, and to do so with at least 10% and upwards of 150% less control effort. These findings illustrate RL's potential for autonomous nuclear reactor control, laying the groundwork for future integration into high-fidelity simulations and experimental validation efforts.

22 - GENERAL STUDIES OF NUCLEAR REACTORS

Towards Immobilized Proton-Coupled Electron Transfer Agents for Electrochemical Carbon Capture from Air and Seawater

Electrochemical CO 2 separation has drawn attention as a promising strategy for using renewable energy to mitigate climate change. Redox-active compounds that undergo proton-coupled electron transfer (PCET) are an impetus for pH-swing-driven CO 2 capture at low energetic costs. However, multiple barriers hinder this technology from maturing, including sensitivity to oxygen and the slow kinetics of CO 2 capture. Here, we use vapor phase chemistry to construct a textile electrode comprising an immobilized PCET agent, poly(1-aminoanthraquinone) (PAAQ), and incorporate it into redox flow cells. This design contrasts with others that use dissolved PCET agents by confining proton-storage to the surface of an electrode kept separate from an aqueous, CO 2 -capturing phase. This system facilitates carbon capture from gaseous sources (a 1% CO 2 feed and air), as well as seawater, with the latter at an energetic cost of 202 kJ/mol CO2 , and we find that quinone moieties embedded within the electrode are more stable to oxygen than dissolved counterparts. Simulations using a 1D reaction-transport model show that moderate energetic costs should be possible for air capture of CO 2 with higher loadings of polymer-bound PCET moieties. The remarkable stability of this system sets the stage for producing textile-based electrodes that facilitate pH-swing-driven carbon capture in practical situations.

Ali, Fawaz

MSD CoP Webinar: "Generative agents: A new frontier for representing human actors and their behavior in MSD models"

Context: This webinar was hosted by the MultiSector Dynamics Community of Practice (MSD CoP; https://multisectordynamics.org). Talk #1: Behavioral Generative Agents for Energy Operations Presenter: Dr. Cong Chen (Thayer School of Engineering, Dartmouth College) Abstract: Accurately modeling consumer behavior in energy operations remains challenging due to inherent uncertainties, behavioral complexities, and limited empirical data. This talk introduces a novel approach leveraging generative agents--artificial agents powered by large language models--to realistically simulate customer decision-making in dynamic energy operations. Talk #2: Simulating multiple human perspectives in socio-ecological systems using large language models Presenter: Dr. Yongchao Zeng (Institute of Meteorology and Climate Research, Atmospheric Environmental Research (IMK-IFU) of the Karlsruhe Institute of Technology in Germany) Abstract: Understanding socio-ecological systems requires insights from diverse stakeholder perspectives. This talk describes a novel simulation system called HoPeS (Human-oriented Perspective Shifting). HoPeS enables model users to not only explore simulated socio-ecological systems (SESs) from a third-person observer's perspective but also take any of the simulated stakeholder roles, like playing an RPG game. By shifting multiple perspectives, model users can reflect and integrate the situated knowledge learned through the participatory simulation, approximating a more holistic and less biased understanding of SESs. Moderators: Jim Yoon (MSD CoP Human Systems Modeling Working Group Co-Chair); Stefano Galelli (MSD CoP Using AI to Enhance MSD Research Working Group Co-Chair); Patrick M. Reed (MSD CoP Facilitation Team) This webinar was held on: November 13th, 2025 from 12-1 PM EST.

Artificial Intelligence

SEAS Communication Engine: An Extensible, Flexible Wrapper for Co-Simulation Agents

When modeling and analyzing the power grid and other large scale systems, researchers often express scenarios as optimization problems and feed them into advanced software solvers. In order to allow multiple solvers to communicate with each other and share data from different domains, the National Renewable Energy Laboratory (NREL) and associated Department of Energy (DOE) labs have developed a software framework called the Hierarchical Engine for Large-scale Infrastructure Co-Simulation (HELICS). HELICS allows cosimulation via a collection of client libraries for different languages that can be called from the appropriate optimization software. However, these client libraries do not provide a higher level of abstraction beyond reading and writing data off of the shared HELICS bus. In this paper, we describe a new software library called the SEAS Communication Engine that exposes a higher-level API for running cosimulation problems. The SEAS Engine provides a class-based abstraction on top of the Python HELICS client, in order to allow users to implement their domain-specific cosimulations without needing to interact with core HELICS primitives. This will make adoption of HELICS and cosimulation in general easier, by exposing a simpler API. In the second part of the paper, we validate our library on a collection of different simulation examples, including the canonical IEEE 13 Bus Feeder. Lastly, we demonstrate using the SEAS Engine to directly call domain-specific code written in the Julia programming language. Our hope is that this will serve as a template for easily calling software in different programming languages via the SEAS Engine, thereby avoiding code duplication and complexity.

co-simulation

Data Assimilation for Robust UQ Within Agent-Based Simulation on HPC Systems

Agent-based simulation provides a powerful tool for in silico system modeling. However, these simulations do not provide built-in methods for uncertainty quantification (UQ). Within these types of models a typical approach to UQ is to run multiple realizations of the model then compute aggregate statistics. This approach is limited due to the compute time required for a solution. When faced with an emerging biothreat, public health decisions need to be made quickly and solutions for integrating near real-time data with analytic tools are needed. We propose an integrated Bayesian UQ framework for agent-based models based on sequential Monte Carlo sampling. Given streaming or static data about the evolution of an emerging pathogen this Bayesian framework provides a distribution over the parameters governing the spread of a disease through a population. These estimates of the spread of a disease may be provided to public health agencies seeking to abate the spread. By coupling agent-based simulations with Bayesian modeling in a data assimilation, our proposed framework provides a powerful tool for modeling dynamical systems in silico. We propose a method which reduces model error and provides a range of realistic possible outcomes. Moreover, our method addresses two primary limitations of ABMs: the lack of UQ and an inability to assimilate data. Our proposed framework combines the flexibility of an agent-based model with UQ provided by the Bayesian paradigm in a workflow which scales well to HPC systems. We provide algorithmic details and results on a simulated outbreak with both static and streaming data.

Spannaus, Adam [ORNL] (ORCID:0000000225213657)

Data for Filling the Cellulosic Bio-economy Gap by Utilizing a Wedge Approach Combined with Stakeholder Collaboration

The price gap between the market and breakeven prices of cellulosic biomass for farmers represents a significant barrier to the development of a low-carbon cellulosic bioeconomy. Using a bottom-up, agent-based modeling tool that replicates the behaviors and interactions of key stakeholders, this study analyzes the emergence of a cellulosic bioeconomy at the local scale through a wedge approach that examines an integrated portfolio of multiple policy options, including subsidies for small-scale bioproducts and environmental credits. The role of collaboration among multiple stakeholders, such as biomass producers (farmers), bio-refinery industry, government, and society, is assessed for filling the price gap. Using the Sangamon River Basin as a case study site, we evaluate the effectiveness of the wedge approach by comparing simulation results from multiple scenarios, each incorporating different combinations of bioeconomy wedges, with and without stakeholder collaboration. Results underscore that active collaboration among stakeholders acts as a catalyst enlarging the effectiveness of bioeconomy wedges. Including the carbon credits and environmental value in the policy portfolio is found to bridge the price gap through collective contributions from diverse stakeholders, where the cellulosic biofuel and bioproduct industry plays a pivotal role. Although this study is conducted at the local watershed scale, the methodology and findings offer valuable insights for market development in other watersheds and the potential scaling of local markets to regional and national levels.

Economics

Co-orchestration of multiple instruments to uncover structure–property relationships in combinatorial libraries

The rapid growth of automated and autonomous instrumentation brings forth opportunities for the co-orchestration of multimodal tools that are equipped with multiple sequential detection methods or several characterization techniques to explore identical samples. This is exemplified by combinatorial libraries that can be explored in multiple locations via multiple tools simultaneously or downstream characterization in automated synthesis systems. In co-orchestration approaches, information gained in one modality should accelerate the discovery of other modalities. Correspondingly, an orchestrating agent should select the measurement modality based on the anticipated knowledge gain and measurement cost. Herein, we propose and implement a co-orchestration approach for conducting measurements with complex observables, such as spectra or images. The method relies on combining dimensionality reduction by variational autoencoders with representation learning for control over the latent space structure and integration into an iterative workflow via multi-task Gaussian Processes (GPs). This approach further allows for the native incorporation of the system's physics via a probabilistic model as a mean function of the GPs. We illustrate this method for different modes of piezoresponse force microscopy and micro-Raman spectroscopy on a combinatorial Sm-BiFeO3 library. However, the proposed framework is general and can be extended to multiple measurement modalities and arbitrary dimensionality of the measured signals.

47 OTHER INSTRUMENTATION

Filling the cellulosic bio-economy gap by utilizing a wedge approach combined with stakeholder collaboration

The price gap between the market and breakeven prices of cellulosic biomass for farmers represents a significant barrier to the development of a low-carbon cellulosic bioeconomy. Using a bottom-up, agent-based modeling tool that replicates the behaviors and interactions of key stakeholders, this study analyzes the emergence of a cellulosic bioeconomy at the local scale through a wedge approach that examines an integrated portfolio of multiple policy options, including subsidies for small-scale bioproducts and environmental credits. Here, the role of collaboration among multiple stakeholders, such as biomass producers (farmers), bio-refinery industry, government, and society, is assessed for filling the price gap. Using the Sangamon River Basin as a case study site, we evaluate the effectiveness of the wedge approach by comparing simulation results from multiple scenarios, each incorporating different combinations of bioeconomy wedges, with and without stakeholder collaboration. Results underscore that active collaboration among stakeholders acts as a catalyst enlarging the effectiveness of bioeconomy wedges. Including the carbon credits and environmental value in the policy portfolio is found to bridge the price gap through collective contributions from diverse stakeholders, where the cellulosic biofuel and bioproduct industry plays a pivotal role. Although this study is conducted at the local watershed scale, the methodology and findings offer valuable insights for market development in other watersheds and the potential scaling of local markets to regional and national levels.

09 BIOMASS FUELS

Harnessing Machine Learning for Agnostic Biodetection

The United States’ current list-based approach to biodefense is limited because it considers only known biological agents. Alternatively, developing and adopting a system based on agent-agnostic signatures would enable detection and characterization of both known and novel agents, thereby engendering greater adaptability in the face of an evolving threat landscape. Machine learning (ML) could aid in such a transition, as it can recognize and encode highly complex patterns from multiple input data modalities and has already demonstrated success in many healthcare and defense applications. Functionalizing ML for environmental biodetection requires understanding current technical capabilities. In this article, we provide a systematic review of existing ML platforms and discuss anticipated development efforts needed to achieve effective ML-enabled, agnostic biodetection.

60 APPLIED LIFE SCIENCES

A Novel LDPP-MADDPG Approach for Distributed Power Allocation in mmWave Cellular Networks

This paper considers the problem of distributed beam scheduling and power allocation problem in millimeter- Wave (mmWave) cellular networks, in which multiple Base Stations (BSs) operate as individual operators over a shared spectrum. We propose a novel learning-aided approach that integrates the Lyapunov Drift-Plus-Penalty (LDPP) framework and Multi-agent Deep Deterministic Policy Gradient (MADDPG) reinforcement learning algorithms. This offers a powerful approach to learning stable and constraint-aware policies, reaping the joint benefit of both LDPP and MADDPG, in complex multiagent environments. The major challenge for this approach is to integrate these two approaches in a meaningful and effective manner. The key idea to solve this problem is to introduce a novel feature of local observation that incorporates potential negative value of the reward function due to the stochastic constraints introduced by the LDPP framework. Empirical results demonstrate that our proposed scheme outperforms the baseline methods under various conditions.

99 - GENERAL AND MISCELLANEOUS