Search NASASearch

SEARCH · Search NASA

Results for “multiple agents”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Topology-Aware Reinforcement Learning for Voltage Control: Centralized and Decentralized Strategies

Volt-VAR control (VVC) methods based on deep reinforcement learning (DRL) can effectively control distribution grid voltage and minimize power loss by implementing corrective and preventive control measures on the reactive power output of inverter-based distributed energy resources (DERs). However, model-free DRL-based VVC approaches usually cannot capture the important topological feature of the power system since they use a fully-connected network (FCN) to deliver the action. Therefore, this paper proposes a graph convolutional network (GCN)-based DRL approach that can employ the topological information of the network to take better control action for regulating the voltage. Our implementation allows for both centralized and decentralized configurations, utilizing a single agent and multiple agents respectively. Although the centralized GCN-based DRL approach has its advantages of minimizing voltage fluctuation and power loss, it is not suitable for large scale power systems due to its challenges in terms of scalability, computation speed and potential single points of failure. Therefore, these problems can be resolved using the decentralized GCN-based DRL approach. Moreover, to ensure the safe operation of the model, our proposed approach incorporates an exponential barrier function while formulating the reward function for each agent. To validate performance of the proposed approaches, the proposed model is tested on modified IEEE test systems and the performances are measured in terms on voltage fluctuation reduction, minimization of power loss and computational speed. Finally, the results show that the proposed topology-aware approach outperforms the FCN-based DRL approach in terms of reducing voltage fluctuation and minimizing power loss of the network. Moreover, it is shown that the decentralized GCN-based DRL has faster computational speed than other approaches.

42 ENGINEERING

Neural correspondence to spectrum of environmental uncertainty in multiple-cue probability judgment system with time delay

Despite state-of-the-art technologies like artificial intelligence, human judgment is critically essential in cooperative systems, such as the multi-agent system (MAS), which collect information among agents based on multiple-cue judgment. Human agents can prevent impaired situational awareness of automated agents by confirming situations under environmental uncertainty. System error caused by uncertainty can result in an unreliable system environment, and this environment affects the human agent, resulting in non-optimal decision-making in MAS. Thus, it is necessary to know how human behavior is changed to capture system reliability under uncertainty. Another issue affecting MAS is time delay, which can delay agent information transfer, resulting in low performance and instability. However, it is difficult to find studies on the influence of time delay on human agents. This study is about understanding the human decision-making process under a specific system reliability environment by uncertainty with time delay. We used concepts of expected and unexpected uncertainty to implement reliability of the system usage environment with three types of time delay conditions: no time delay, regular time delay, and irregular time delay conditions. We used electroencephalogram (EEG) for human cognitive neural mechanisms in multiple-cue judgment systems to understand human decision-making. In the reliability of system usage environment, the unreliable system environment significantly creates less memory load by less utilization of system rules for decision-making. In terms of time delay, delayed information delivery does not significantly affect memory load for decision-making.

cognitive process

A multi-scale cognitive interaction model of instrument operations at the Linac Coherent Light Source

The Linac Coherent Light Source (LCLS) is the world’s first x-ray free electron laser. It is a scientific user facility operated by the SLAC National Accelerator Laboratory, at Stanford, for the U.S. Department of Energy. As beam time at LCLS is extremely valuable and limited, experimental efficiency—getting the most high quality data in the least time—is critical. Our overall project employs cognitive engineering methodologies with the goal of improving experimental efficiency and increasing scientific productivity at LCLS by refining experimental interfaces and workflows, simplifying tasks, reducing errors, and improving operator safety and stress. Here, in this study, we describe a multi-agent, multi-scale computational cognitive interaction model of instrument operations at LCLS. Our model simulates the aspects of human cognition at multiple cognitive and temporal scales, ranging from seconds to hours, and among agents playing multiple roles, including instrument operator, real time data analyst, and experiment manager. The model can roughly predict impacts stemming from proposed changes to operational interfaces and workflows. Example results demonstrate the model’s potential in guiding modifications to improve operational efficiency. We discuss the implications of our effort for cognitive engineering in complex experimental settings and outline future directions for research. The model is open source, and the videos of the supplementary material provide extensive detail.

47 OTHER INSTRUMENTATION

Analysis of strategies to meet ASHRAE S241 infectious aerosol control targets by space type and region using EnergyPlus™ simulations

Professional organizations, such as ASHRAE, have recently proposed new voluntary standards for controlling infectious aerosols. Specifically, ASHRAE Standard 241 (S241) defines equivalent clean air targets for a range of space types that can be met through combinations of mitigation measures. This paper seeks to inform the selection of measures by space type and climate zone through building simulations that quantify the impacts of increased outdoor air ventilation; increasing filtration and/or adding germicidal ultraviolet (GUV) in the central heating, ventilation, and air-conditioning (HVAC) systems; or using portable air cleaners (PACs), upper room GUV, or whole room GUV approaches to increase equivalent clean air delivery. The measures are assessed individually and in practical combinations for their ability to meet S241 in four space types (offices, classrooms, dining areas, and healthcare waiting rooms) using prototype buildings modeled using EnergyPlus. The mitigation measures are compared holistically against baseline building operations using metrics for equivalent clean air, energy, and comfort. The results in this study show that single measures can meet S241 for offices with minimal impacts on energy and comfort, while either upper room GUV systems or combinations of measures such as MERV 13 HVAC filtration with PACs are needed to meet S241 for classrooms. The dining area and healthcare waiting room targets cannot be met in this study when assuming design occupancy and MS2 as the challenge agent, even when combining multiple measures together. This paper provides valuable considerations when designing measures to meet S241 for a range of spaces and scenarios.

GUV

Deep reinforcement learning control for co-optimizing energy consumption, thermal comfort, and indoor air quality in an office building

With the recent demand for decarbonization and energy efficiency, advanced HVAC control using Deep Reinforcement Learning (DRL) becomes a promising solution. Due to its flexible structures, DRL has been successful in energy reduction for many HVAC systems. However, only a few researches applied DRL agents to manage the entire central HVAC system and control multiple components in both the water loop and the air loop, owing to its complex system structures. Moreover, those researches have not extended their applications by incorporating the indoor air quality, especially both CO2 and PM2.5concentrations, on top of energy saving and thermal comfort, as achieving those objectives simultaneously can cause multiple control conflicts. What's more, DRL agents are usually trained on the simulation environment before deployment, so another challenge is to develop an accurate but relatively simple simulator. Therefore, we propose a DRL algorithm for a central HVAC system to co-optimize energy consumption, thermal comfort, indoor CO2 level, and indoor PM2.5 level in an office building. To train the controller, we also developed a hybrid simulator that decoupled the complex system into multiple simulation models, which are calibrated separately using laboratory test data. The hybrid simulator combined the dynamics of the HVAC system, the building envelope, as well as moisture, CO2, and particulate matter transfer. Three control algorithms (rule-based, MPC, and DRL) are developed, and their performances are evaluated on the hybrid simulator environment with a realistic scenario (i.e., with stochastic noises). The test results showed that, the DRL controller can save 21.4 % of energy compared to a rule-based controller, and has improved thermal comfort, reduced indoor CO2 concentration. The MPC controller showed an 18.6 % energy saving compared to the DRL controller, mainly due to savings from comfort and indoor air quality boundary violations caused by unmeasured disturbances, and it also highlights computational challenges in real-time control due to non-linear optimization. Finally, we provide the practical considerations for designing and implementing the DRL and MPC controllers based on their respective pros and cons.

Guo, Fangzhou

Precipitation of gadolinium from magnetic resonance imaging contrast agents may be the Brass tacks of toxicity

The formation of gadolinium-rich nanoparticles in multiple tissues from intravenous magnetic resonance imaging contrast agents may be the initial step in rare earth metallosis. The mechanism of gadolinium-induced diseases is poorly understood, as is how these characteristic nanoparticles are formed. Gadolinium deposition has been observed with all magnetic resonance imaging contrast agent brands. Aside from endogenous metals and acidic conditions, little attention has been paid to the role of the biological milieu in the degradation of magnetic resonance imaging contrast agents into nanoparticles. Herein, we describe the decomposition of the commercial magnetic resonance imaging contrast agents Omniscan and Dotarem in the presence of oxalic acid, a well-known endogenous compound. Omniscan dechelated rapidly and preluded measurement by the means available, while Dotarem underwent a two-step decomposition process. The decomposition of both magnetic resonance imaging contrast agents by oxalic acid formed gadolinium oxalate (Gd 2 [C 2 O 4 ] 3 , Gd 2 Ox 3 ). Furthermore, both observed steps of the Dotarem reaction involved the associative addition of oxalic acid. Adding protein (bovine serum albumin) increased the rate of dechelation. Displacement reactions could occur at lysosomal pH. Through these studies, we have demonstrated that magnetic resonance imaging contrast agents can be dissociated by endogenous molecules, thus illustrating a metric by which gadolinium-based contrast agents (GBCAs) might be destabilized in vivo.

Gadodiamide

Wholesale Electricity Market Design to Support Resource Adequacy

Wholesale electricity markets are intended to incentivize system generation investments and operations outcomes that meet evolving system needs. In this work, we evaluate the effectiveness of wholesale market structures, rules and policies in achieving system resource adequacy (RA) and clean energy targets in the presence of self-interested generation investors using the Electricity Markets and Investment Suite Agent-based Simulation (EMIS-AS) model. Results highlight that both capacity markets and operating reserve demand curves (ORDCs) can help achieve a reliable system but with different RA compliance timelines and distribution of generation technologies. Structures with capacity markets tend to favor more capital-intensive peaking technologies while reducing wind and solar build-outs due to suppressed energy and clean energy market prices, particularly in the absence of strong clean energy targets. Conversely, ORDCs improve the commitment of available generation units, but this comes at the expense of higher system costs and renewable generation curtailment. We also find that well-calibrated static capacity demand curves can yield similar reliability and total cost compared to capacity market demand curves informed dynamically by resource adequacy while also yielding stable annual capacity prices. Different approaches to formulating ORDC curves can also yield key trade-offs, namely that a more efficient treatment of storage chronology results in lower ORDC curves and prices, yielding less investment and cost but at the expense of reliability. Finally, the effectiveness of wholesale electricity markets in practically achieving very high clean energy generation targets highly depends on the cost-competitiveness of clean energy technologies that can support critical balancing needs across multiple timescales.

agent based modeling

Entanglement engineering of optomechanical systems by reinforcement learning

Entanglement is fundamental to quantum information science and technology, yet controlling and manipulating entanglement—so-called entanglement engineering—for arbitrary quantum systems remains a formidable challenge. There are two difficulties: the fragility of quantum entanglement and its experimental characterization. We develop a model-free deep reinforcement-learning (RL) approach to entanglement engineering, in which feedback control together with weak continuous measurement and partial state observation is exploited to generate and maintain desired entanglement. We employ quantum optomechanical systems with linear or nonlinear photon–phonon interactions to demonstrate the workings of our machine-learning-based entanglement engineering protocol. In particular, the RL agent sequentially interacts with one or multiple parallel quantum optomechanical environments, collects trajectories, and updates the policy to maximize the accumulated reward to create and stabilize quantum entanglement over an arbitrary amount of time. The machine-learning-based model-free control principle is applicable to the entanglement engineering of experimental quantum systems in general.

97 MATHEMATICS AND COMPUTING

Scalable Fabrication of a Fibrous Amine-functionalized Matrix (FAM) Sorbent for Critical Mineral Recovery

We report a novel flat sheet Fibrous Amine-functionalized Matrix (FAM) sorbent platform designed for efficient and selective capture of CM from dilute solutions. The FAM sorbent features crosslinked amine microfilms coated onto/within a glass fiber matrix, providing fast mass transfer and excellent mechanical stability. Systematic batch and flow-through tests with FAM revealed rapid metal uptake kinetics and high capacity for representative species, achieving ~90 mg/g of Gallium, ~100 mg/g of Cobalt, and ~90 mg/g for Neodymium. Moreover, multiple eluents, including mineral acids and complexing agents, enabled highly effective desorption of adsorbed metals, demonstrating the feasibility of regenerating FAM sorbents. Importantly, tests with authentic coal ash leachate demonstrated strong selectivity toward U.S. Department of Energy (DOE)-listed CM and rare earth elements over abundant base cations, confirming the robustness of FAM in realistic complex solutions. The flat sheet geometry was amenable to scaling into durable spiral wound modules, highlighting the potential for future regeneration and reuse. This work establishes FAM sorbents as a promising platform for the recovery of CM from wastewaters, advancing both resource sustainability and environmental stewardship.

critical mineral recovery

Agentic traffic intelligence: Augmented human-in-the-loop scenario generation for microscopic traffic simulation

Traditional microscopic traffic simulation generation often relies on static datasets and manual design, limiting its ability to simulate complex conditions easily. This paper presents a novel framework, Agentic Traffic Intelligence, which combines human approval large language models (LLMs), the Real-Twin tool, and multi-agent systems to perform realistic microscopic traffic simulation scenario generation. The proposed framework incorporates human-in-the-loop (HIL) control, retrieval-augmented generation (RAG), and multi-agent control mechanisms. HIL mechanisms are used to guide multiple LLMs focused on attributes for microscopic simulation generation and to improve the interpretability and transparency of LLM execution for users. RAG enhances context extraction by dynamically integrating external knowledge sources for traffic scenario generation foundations. A multi-agent architecture with supervisory control coordinates the interaction of simulation components, including traffic simulators, control logic, and calibration tools. This enables the synthesis of simulation-ready scenarios that reflect dynamic demand profiles and behavior controls. Furthermore, the framework fuses multisource traffic data with unstructured context and supports iterative refinement through interactive user feedback. Validated through microscopic simulation using Simulation of Urban Mobility, the generated scenarios demonstrate high-fidelity network generation with inflow and turn movement and behavioral calibration, offering a robust and efficient tool for stress-testing and optimizing urban mobility systems.

Hierarchical multi-agent control

The ballad of LLM agents: philosophical reasoning for chemistry

Large language models (LLMs) show remarkable potential for scientific reasoning but often produce unreliable or scientifically unactionable outputs when faced with multi-step logic, domain grounding, and interpretability challenges, especially in complex fields like chemistry and materials science. Here, we introduce a framework of philosophical reasoning agents, inspired by canonical thinkers such as Socrates, Descartes, Kant, and Hume, to guide LLM behavior via structured prompt engineering. These agents embody distinct reasoning paradigms (dialectical inquiry, deductive logic, rule-based judgment, and empirical validation) and are evaluated across multiple chemistry subdomains, physical, analytical, general, inorganic, and organic chemistry, using the ChemBench benchmark. Our agentic prompting approach yields substantial accuracy gains on open-ended numerical chemistry questions, with gains of +11.5 percentage points for GPT-4o with Hume, +4.5 percentage points for GPT-5 with Kant, and +21.8 percentage points for GPT-5.1 with Socrates at the strict 1% error threshold, relative to the corresponding base models. Beyond accuracy, we observe benchmark-level model–agent performance patterns, suggesting that different prompting styles interact differently with each base model. These findings demonstrate that embedding philosophy-of-science principles into multi-agent frameworks can improve and produce interpretable, adaptive, and domain-aligned scientific LLMs.

Harb, Hassan [Argonne National Laboratory (ANL), A

Thermodynamically leveraged solventless aerobic deconstruction of polyethylene-terephthalate plastics over a single-site molybdenum-dioxo catalyst

Here, we describe the solventless catalytic deconstruction of polyethylene-terephthalate (PET) under an aerobic atmosphere, mediated by an earth-abundant, low-cost activated carbon (AC)-supported single-site molybdenum-dioxo catalyst (AC/MoO 2 ). Catalytic amounts of AC/MoO 2 selectively convert waste PET into its monomer, terephthalic acid (TPA), within 4 h at 265 °C with yields as high as 94% under 1 atm air. Pure crystalline TPA product sublimes from the reaction hot zone, crystallizing on the reactor cold zone, thus avoiding the need for separation and purification steps. This process does not employ any hazardous/toxic reducing agents or solvents, and the catalyst can be recycled multiple times without loss of activity, rendering this process highly atom-efficient. According to computational and experimental mechanistic studies, the AC/MoO 2 catalyst mediates a thermoneutral metal-catalyzed β-scission step, followed by an exothermic step that converts the vinyl benzoate intermediate to TPA and acetaldehyde using trace amounts of moisture in the air. The formation of gaseous acetaldehyde makes the isolation of TPA from the reaction mixture facile and industrially favorable, especially since solvents are unnecessary. The present methodology is also extended to the deconstruction of other frequently used polyester plastics, polybutylene terephthalate (PBT), polyethylene naphthalate (PEN), and polyethylene furanoate (PEF), and operates equally well with post-consumer waste products. Notably, this process is also compatible with plastic mixtures of polyesters with polyolefins, polyamides, and polycarbonates, leading to the selective conversion of each polyester to the corresponding monomer, leaving the residual polymer unchanged and polyester-free.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

One-Pot Self-Assembly of Sequence-Controlled Mesoporous Heterostructures via Structure-Directing Agents

Multimaterial heterostructures have led to characteristics surpassing the individual components. Nature controls the architecture and placement of multiple materials through biomineralization of nanoparticles (NPs); however, synthetic heterostructure formation remains limited and generally departs from the elegance of self-assembly. Here, in this study, a class of block polymer structure-directing agents (SDAs) are developed containing repeat units capable of persistent (covalent) NP interactions that enable the direct fabrication of nanoscale porous heterostructures, where a single material is localized at the pore surface as a continuous layer. This SDA binding motif (design rule 1) enables sequence-controlled heterostructures, where the composition profile and interfaces correspond to the synthetic addition order. This approach is generalized with 5 material sequences using an SDA with only persistent SDA-NP interactions (“P-NP 1 –NP 2 ”; NP i = TiO 2 , Nb 2 O 5 , ZrO 2 ). Expanding these polymer SDA design guidelines, it is shown that the combination of both persistent and dynamic (noncovalent) SDA-NP interactions (“PD-NP 1 –NP 2 ”) improves the production of uniform interconnected porosity (design rule 2). The resulting competitive binding between two segments of the SDA (P- vs D-) requires additional time for the first NP type (NP 1 ) to reach and covalently attach to the SDA (design rule 3). The combination of these three design rules enables the direct self-assembly of heterostructures that localize a single material at the pore surface while preserving continuous porosity.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Distillable amine-based solvents for effective pretreatment of multiple biomass feedstocks

Exploring the potential of advanced distillable solvents as efficient biomass pretreatment agents is critical for biorefineries, enhancing fermentable sugar yields while enabling solvent recovery and recycling without suffering significant losses. Here, we employ distillable amine-based solvents for pretreating a wide range of lignocellulosic feedstocks, aiming to facilitate the industrial release of fermentable sugars from diverse feedstocks through enzymatic hydrolysis. Twenty-two diverse feedstocks, sourced from different geographical regions and representing various biomass categories, were surveyed for chemical (mainly carbohydrates and lignin) and lignin (S, G, and H units) profiles. Several solvents, including ethanolamine, ethanolammonium acetate, butylamine, butylammonium acetate, and triethylamine, were tested for the pretreatment of eight selected biomasses. Among these solvents, butylamine emerged as the most effective due to its favorable sugar release, excellent solvent removal rate, and low boiling point, facilitating solvent recovery and recycling. Extending butylamine pretreatment to all 22 feedstocks demonstrated desirable sugar yields and highly efficient solvent removal in the majority of the biomass sources tested. Agricultural residues and their mixtures showed particularly favorable sugar release. Despite minimal changes in cellulose crystallinity, XRD characterization of sorghum, poplar, and pine before and after butylamine pretreatment showed a decrease in intensity and a slight shift of certain peaks, indicating alterations in cellulose structure. Fourier-transform infrared spectroscopy and thermogravimetric analysis analyses suggested disruption of biomass linkages in hemicellulose and lignin, enhancing enzymatic digestibility. Scale-up experiments of the mixed agricultural feedstocks in a 1 L Parr reactor achieved over 90% glucose liberation and more than 99% butylamine removal, highlighting the scalability of the method. The resulting hydrolysates supported the growth of diverse bacterial and fungal strains, indicating downstream compatibility with commercial fermentation processes. This study presents butylamine as an effective, recoverable pretreatment solvent for a wide range of lignocellulosic feedstocks, offering a promising solution to key biorefinery challenges. The demonstrated scalability and compatibility with various biomass types and blends underscore its potential for industrial application, advancing sustainable biofuel and biochemical production.

biomass composition

Developing Fluorescence-Based Sensors to Support Rare Earth Element Separation

Rare earth elements (REEs) are essential to most renewable energy technologies. Unfortunately, as we transition to sustainable energy production, the demand for REEs is rapidly growing well beyond current rates of production. As a result, novel means of efficient, scalable, and easily adaptable methods for processing primary and recycle feedstocks are needed. Development and integration of sensors for highly selective in-line monitoring can support more efficient design and testing of such novel separation processes, as well as more cost-effective deployment of those separation flowsheets. Work here will explore the application of fluorescence spectroscopy, a highly sensitive and selective technique, to quantify multiple lanthanides in complex mixtures including known interferents or quenching agents. Results include identification of the optimal excitation wavelength and the limit of detection of various rare earth elements as well as the performance of data-science-based quantification approaches in streams where “unknowns” are present. Overall, the data science tools in conjunction with optical sensor data were able to quantify analytes in the presence of other lanthanides which can be anticipated in the actual industrial stream. Here we include characterization of lanthanides in a microfluidic device similar to those used in new process development. This study demonstrates the capability of utilizing fluorescence spectroscopy to quantify analytes in a complicated solution matrix, suggesting this is a successful approach for in-line monitoring to optimize the separation efficiency in an industrial stream.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Design, Preparation, and Execution of the 100-AV Field Test for the CIRCLES Consortium: Methodology and Implementation of the Largest Mobile Traffic Control Experiment to Date

This article presents the comprehensive design, setup, execution, and evaluation of the MegaVanderTest (MVT) experiment conducted by the Congestion Impacts Reduction via CAV-in-the-Loop Lagrangian Energy Smoothing (CIRCLES) Consortium, which aimed to mitigate traffic congestion using partially autonomous vehicles (AVs) (see “Summary”). The experiment involved 100 vehicles on Nashville’s Interstate 24 (I-24) highway, utilizing various control algorithms to smooth stop-and-go traffic waves. The execution of the MVT experiment required a coordinated effort from multiple teams. This article details the meticulous planning process, the coordinated efforts of multiple teams, and the innovative use of a dynamic agent-based simulation framework for traffic evaluation. Here, the contributions of this work include demonstrating and providing a detailed roadmap for large-scale live traffic experiments, illustrating the lessons learned from the MVT experiment, and introducing the other articles in this issue and their complementary relationship in the MVT experiment.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI

Adaptive Reinforcement Learning Control for Power Distribution in Multi-Output Resonant Converters

This paper presents an adaptive reinforcement learning (ARL)-based control framework for efficient power distribution in a multi-output resonant converter for UAV applications. The proposed system is based on a high-frequency isolated resonant architecture, where a single energy source supplies multiple propulsion loads through independently controlled output rectifiers, addressing the need for coordinated multi-motor power management. The ARL framework dynamically allocates output power by learning optimal phase-shift control actions under varying load demands and operating conditions. The agent autonomously determines control parameters that maximize conversion efficiency while ensuring accurate power sharing among multiple outputs. In addition, the proposed approach enables adaptive operation without requiring detailed system modeling or manual tuning. Experimental results demonstrate stable and efficient performance over a wide range of operating conditions, confirming the effectiveness and robustness of the learning-based control strategy for multi-output resonant converter system.

Asa, Erdem [ORNL] (ORCID:0000000190884812)