Search NASASearch

SEARCH · Search NASA

Results for “Autonomous Control Systems”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Practical Insights on Applying Simulation-Based Control Methods in Experimental Studies

Advanced nuclear reactors are crucial to the future of energy both in the United States and around the globe. In contrast to the current operating fleet, they are characterized as being deployable in remote locations and able to operate in semi-autonomous or autonomous fashion. This leap forward necessitates a new reactor control paradigm. Because advanced nuclear reactors are still under development in the United States, the creation of new control methods to achieve autonomous operations has been based on systems modeling and simulation. However, an important factor in successfully deploying these new control methods is the ability to seamlessly transition from simulation environments to real-world settings. Control methods tested in both simulation and experimental settings need to be investigated in the context of advanced reactor applications. This work developed a series of simple controllers for Idaho National Laboratory (INL)’s Microreactor Applications Research Validation and Evaluation (MARVEL) microreactor operating in load-following scenarios. These controllers were tested in both simulation and experimental settings, and a comparative performance analysis was performed. The simulation tests leveraged the Control and Optimization Modular Modeling Application for Nuclear Deployment (COMMAND) software developed in a previous stage of the current effort, along with the MARVEL Reactor Excursion and Leak Analysis Program (RELAP5-3D) and Monte Carlo N-Particle (MCNP) models. The experimental tests leveraged the COMMAND software, MARVEL models, and the U.S. Department of Energy Microreactor Program’s Microreactor Automated Control System (MACS). MACS was developed to serve as a control method testbed. It was customized to mirror the MARVEL microreactor, and COMMAND enabled MACS to emulate the physics of MARVEL. The load-following controller was developed using the simulation platform, with efforts to emulate real systems by introducing actuator saturation and noise. These factors were incrementally accounted for in the controller design. After finalizing the controller design, it was implemented with the experimental setup. The experimental conditions tested included an initial test under conditions similar to the final simulation test, and two additional scenarios. The first scenario introduced additional actuator saturation to account for equipment aging over time, which was unknown to the controller. The second scenario introduced sensor delay, a phenomenon anticipated with the use of remote operations or wireless communication in advanced reactors. These tests revealed several notable differences. While the controller performed well in simulation, it exhibited several limitations when transitioning to hardware. The main challenges involved maintaining the steady-state target power, as evidenced by larger error values between the true reactor power and setpoint power, as well as persistent oscillations in controlled reactor power. These issues could lead to unacceptable transient conditions in real reactor testing. Introducing actuator aging and stochastic delays in the experimental setup significantly impacted controller performance, resulting in increased overshoot and undershoot, and exacerbated error and oscillations previously mentioned. These findings underscore the importance of experimental testbeds for testing and validating control methods, as controllers developed using only theory and/or simulation may perform unexpectedly when applied to actual hardware. This research emphasizes the need for an experimental testbed for achieving such validation.

46 - INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AN

The ReSWARM microgravity flight experiments: Planning, control, and model estimation for on‐orbit close proximity operations

Abstract On‐orbit close proximity operations involve robotic spacecraft maneuvering and making decisions for a growing number of mission scenarios demanding autonomy, including on‐orbit assembly, repair, and astronaut assistance. Of these scenarios, on‐orbit assembly is an enabling technology that will allow large space structures to be built in situ, using smaller building block modules. However, like many of these scenarios, robotic on‐orbit assembly involves several technical hurdles, such as changing system models. For instance, grappled modules moved by a free‐flying “assembler” robot can cause significant changes in the combined system inertia, which have cascading impacts on motion planning and control portions of the autonomy stack. Further, on‐orbit assembly and other scenarios require collision‐avoiding motion planning, particularly when operating in a “construction site” scenario of multiple assembler robots and structures. Multiple key technologies that address these complicating factors for autonomous microgravity close proximity operations are detailed in this work, in particular: (1) application of global long‐horizon planning, accomplished using offline and online sampling‐based planner options that consider the system dynamics; (2) adaptation of the recently proposed RATTLE information‐aware planning framework for on‐orbit reconfiguration model learning; and (3) connection with robust control tools to provide low‐level control robustness using current system knowledge. These approaches were demonstrated for an autonomous on‐orbit assembly use case by the RElative Satellite sWarming and Robotic Maneuvering (ReSWARM) experiments using NASA's Astrobee robots on the International Space Station. Results of the ReSWARM experiments are provided along with significant operational and implementation detail discussing the practicalities of hardware implementation and unique aspects of working with the Astrobee free‐flyer robots in microgravity. ReSWARM provides a base set of planning and control tools for robotic close proximity operations, demonstrates them in microgravity, and outlines some of the important hardware aspects that future autonomous free‐flyers will need to consider.

Robotics

Uncertainty quantification and sensitivity analysis of a nuclear thermal propulsion reactor startup sequence

The research presented in this article describes progress in applying stochastic methods, uncertainty quantification, parametric studies, and variance-based sensitivity analysis (also known as Sobol sensitivity analysis) to a full-core model of a nuclear thermal propulsion (NTP) system simulated via the radiation transport code Griffin to simulate neutronics. Our goal is to develop a reduced-order (surrogate) model that can be rapidly sampled with perturbations to multiple input parameters. In this NTP system, reactivity and power feedback affect the rotation of control drums (CDs), which is itself controlled by a hybrid proportional-integral-derivative (PID) controller actuated by the power demand and reactivity feedback from the numerical model. This model uses reactor kinetic feedback (mean generation time [Λ] and effective delayed neutron fraction [ β eff ] from a transient Griffin simulation executed via Griffin’s improved quasi-static solver to provide the kinetic parameters) as inputs to functions that control the CD rotation angle. By investigating numerous stochastic approaches, we developed a dual-purpose surrogate model of the NTP system, using polynomial regression in the Multiphysics Object-Oriented Simulation Environment (MOOSE) Stochastic Tools Module (STM). The trained model can be rapidly sampled while simultaneously perturbing various input parameters, such as coefficients on the PID control or temperature (directly affecting the neutron cross section). The surrogate model delivers accurate (within 5%) results at speeds orders of magnitude faster (minutes, not days of computational time) than the base model. Once the surrogate model has been trained, distributions of the uncertain parameters can be changed at will to investigate the effects of perturbing multiple inputs as well as the effects of these inputs on the model output. For example, coefficients used in the PID control system may vary due to some type of physical interference, or uncertainty may exist in the temperature of the neutron cross sections in various regions of the reactor. A distribution can be placed on these parameters, and operational boundaries can be determined. The goal of this work is to support development of an advanced control system for operating CDs in a functioning NTP system. This work is a scoping study of the MOOSE STM.

21 - SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLAN

Modeling for a Digital Twin-Based Remote Operation System Framework

New reactor designs and technologies are being developed to grow and advance the nuclear industry. Microreactors are one of the many new concepts for advancing the industry. Microreactors are very small reactors generally designed to have an operating power of 20 MWth or less. They are ideal for many applications in which it would not be feasible to have a large-scale reactor, such as powering remote communities, military bases, and mining sites. Many of these applications currently rely on diesel generators for power, and replacing those generators with the carbon-free energy of a microreactor is a major driving factor for microreactor development. However, most of the use cases for microreactors are in isolated locations where construction and labor costs are much higher. Microreactors will need to be comparable to other energy production methods for their deployment to be successful. Remotely operating the microreactor has a great potential to benefit the economics and make it more cost competitive. With remote operations, the operation facility location could be strategically chosen based on factors such as construction costs, workforce size, etc. The benefits of remote operations could be leveraged even more if the remote operation system is semi-autonomous. If the remote operation system is semi-autonomous, more microreactors could be operated and monitored from one remote operation facility. Additionally, with the system handling some tasks for the human operator, it could reduce the number of operating staff necessary for the microreactor. Digital twins can be used to introduce a level of automation to the remote operation system. Digital twins are capable of component monitoring, system operation and control, and predictive performance [1]. All of those features are important for a successful semi-autonomous remote operation system.

22 GENERAL STUDIES OF NUCLEAR REACTORS

Proof-of-Concept for Sensor Modeling in MOOSE for the Design of Autonomous Nuclear Reactor Control

Autonomous operation is essential for the deployment of microreactors and fission batteries, both in terrestrial and space applications. However, prototypes of microreactors and fission batteries do not exist yet, and even the design space has not been narrowed down conclusively, making the instrumentation and control system design difficult. For this reason, there is a need for flexible computational capabilities to create a numerical stand-in of potential microreactor and fission battery designs. The latter can be used to design and test control strategies to support autonomous operations. In this paper, we describe the initial implementation of a pluggable sensor system for the easy implementation of realistic sensor models in the multiphysics object-oriented simulation environment (MOOSE) framework. This new capability will enable MOOSE users to create a numerical stand-in of microreactors and fission batteries, ultimately allowing them to easily test new control algorithms, and instrumentation strategies for advanced systems in the design phase

22 - GENERAL STUDIES OF NUCLEAR REACTORS

Explainable discrepancy checker and diagnosis for digital Twin-based supervisory control system

By virtually representing a physical object and process, a digital twin (DT) enables optimal autonomous operations by combining classical and novel frameworks in sensors, state predictions, and multi-input/multi-output systems. A DT’s values depend on how well models estimate quantities of interest and on how uncertainty is handled. Moreover, DTs often combine physics-based and data-driven models with mixed fidelities, where classical uncertainty quantification (UQ) struggles with many sources of uncertainty and real-time constraints. Here, this work presents a UQ-based discrepancy checking and diagnosis tool for a DT-based supervisory control system. The tool is developed using metadata from an automated DT development process to learn correlations between sources of uncertainties and outcomes. During operation, it compares predictions with measurements, attributes discrepancies to dominant sources, and recommends parameter and configuration updates. We verify the workflow on a synthetic temperature-control problem and deploy it on a virtual Thermal Energy Delivery System, reducing mismatch and improving control robustness.

22 - GENERAL STUDIES OF NUCLEAR REACTORS

Proof-of-Concept for Sensor Modeling in MOOSE for the Design of Autonomous Nuclear Reactor Control

Autonomous operation is essential for the deployment of microreactors and fission batteries, both in terrestrial and space applications. For this reason, recent studies have investigated autonomous control by using adaptive model predictive control and multi-objective optimization for heat pipe–cooled microreactors under normal and heat pipe failure conditions. However, prototypes of microreactors and fission batteries do not exist yet, and even the design space has not been narrowed down conclusively, making the instrumentation and control system design difficult. For this reason, there is a need for flexible computational capabilities to create a numerical stand-in of potential microreactor and fission battery designs. The latter can be used to design and test control strategies to support autonomous operations. In this poster, we describe the initial implementation of a pluggable sensor system for the easy implementation of realistic sensor models in the multiphysics object-oriented simulation environment (MOOSE) framework. This new capability will enable MOOSE users to create a numerical stand-in of microreactors and fission batteries, ultimately allowing them to easily test new control algorithms, and instrumentation strategies for advanced systems in the design phase.

22 - GENERAL STUDIES OF NUCLEAR REACTORS

A Large-Scale Analysis to Optimize the Control and V2V Communication Protocols for CDA Agreement-Seeking Cooperation

Cooperative driving automation (CDA) Class C, agreement-seeking cooperation, is an innovative and practical solution that can promote cooperation among general passenger vehicles on the road. However, more comprehensive studies are needed before establishing the standard protocols of agreement-seeking cooperation, such as communication frequency and the duration of cooperation. Here, this article presents an initiative study on the impacts of communication capabilities on agreement-seeking cooperation. Through a large-scale analysis by regulating vehicle-to-vehicle (V2V) communication metrics, this work suggests desirable system parameters that can maximize the benefits of cooperation and ensure reliable operability while avoiding exhaustive communication loads. As the first step, an example agreement-seeking cooperation system is created for a car-following scenario, including decision-making and control algorithms for autonomous vehicles. Then, software-in-the-loop tests explore the performance of the developed system as it encounters various communication risks, such as latency and message packet drops. The system performance metrics are evaluated from various angles, including the time consumed for the agreement-seeking process, cooperation ratio, and the ratio of faulty cooperation. Energy saving from the cooperation is assessed by using simulation software that can run multiple high-fidelity vehicle models simultaneously. Based on the analyses, this article suggests the V2V communication requirements for the reliable operation of CDA agreement-seeking, which can be referred to when developing the standard protocols of agreement-seeking cooperation.

42 ENGINEERING

Towards Agentic AI on Particle Accelerators

As particle accelerators grow in complexity, traditional control methods face increasing challenges in achieving optimal performance. This paper envisions a paradigm shift: a decentralized multi-agent framework for accelerator control, powered by Large Language Models (LLMs) and distributed among autonomous agents. We present a proposition of a self-improving decentralized system where intelligent agents handle high-level tasks and communication and each agent is specialized control individual accelerator components. This approach raises some questions: What are the future applications of AI in particle accelerators? How can we implement an autonomous complex system such as a particle accelerator where agents gradually improve through experience and human feedback? What are the implications of integrating a human-in-the-loop component for labeling operational data and providing expert guidance? We show two examples, where we demonstrate viability of such architecture.

43 PARTICLE ACCELERATORS

Unified Universal Control and Coordination of Inverter-Based Resources, and Validation for a PV + Battery Hybrid Plant

As renewable energy deployment grows, hybrid power plants (HPPs) combining photovoltaic (PV) and battery systems must evolve to offer both energy and grid stability services. These systems typically include a mix of grid-following (GFL) and grid-forming (GFM) inverters, presenting unique coordination and control challenges. This Department of Energy–funded project developed and validated a Unified Universal Control and Coordination (UUCC) framework for such PV + battery hybrid plants, enabling seamless and stable operation, including ultrafast black start, autonomous synchronization, and robust frequency and voltage regulation, under different grid conditions. The project significantly advanced the understanding of inverter-based resource (IBR) control by developing and validating three complementary system-level approaches for hybrid GFL/GFM operation: 1. A combined Virtual Resistance (VR)-based GFL and Virtual Oscillator Control (VOC)-based GFM method, where each inverter type is governed by a specialized control strategy. Together, these achieve stable, fast-response coordination, eliminating inrush current and enabling smooth black start and grid synchronization across a wide range of grid strengths. 2. A Deadbeat-based UUCC strategy, which uses discrete-time, switching-cycle-level control for both GFL and GFM inverters. This approach replaces traditional PI/PLL control with a control parameter-free, high-bandwidth framework that supports stable LVRT and instantaneous synchronization under all conditions. 3. A benchmark comparison with Siemens’ commercial GFM microgrid controller, which provided a fast baseline platform. The commercial approach decoupled v & f control was implemented on a commercial microgrid controller.The baseline commercial benchmark helped highlight superior transient response and black start performance offered by the deadbeat and VOC approaches. These technical contributions offer substantial improvements over conventional inverter control schemes, which often rely on slow phase-locked loop (PLL)-based synchronization, require careful control parameters tuning, and prone to unstable in weak grids with GFL inverters and in stiff grid with GFM inverters therefore challenging for hybrid GFL+GFM under all grid conditions. The deadbeat-based UUCC framework enables simpler, faster, and more robust operation of hybrid IBR systems using wide-bandgap (WBG) devices such as SiC power semiconductors. The rapid expansion of hybrid distributed energy resources (DERs), including residential and commercial PV-BESS installations such as Tesla Powerwall, PV with vehicle-to-grid (V2G) capability, and other integrated configurations, presents complex operational challenges for medium-voltage radial distribution feeders. These networks are subject to frequent disturbances such as faults, switching operations, rapid reclosing sequences, and feeder reconfigurations, all of which introduce dynamic stress on IBRs. In addition, planned feeder segmentation and deliberate islanding for resilience will require DERs that can autonomously perform blackstart, establish voltage and frequency references, and resynchronize with the main grid. The advanced deadbeat-based UUCC control and blackstart functionalities developed in this project directly address these requirements, enabling decentralized and autonomous operation of inverter-dominated DERs in distribution systems under a wide range of fault and reconfiguration scenarios. From a public benefit perspective, these innovations enable more reliable and cost-effective integration of renewable energy into distribution networks. The ability to autonomously black start and stabilize grids under varying grid conditions support accelerates recovery from outages and support decentralized resilient energy systems. By reducing system complexity and improving performance, this project lays critical groundwork for future inverter-dominated power grids that are clean, reliable, and accessible to all.

14 SOLAR ENERGY

An Approach to Realize Generalized Optimal Motion Primitives Using Physics Informed Neural Networks

Autonomous manipulation is a challenging problem in field robotics due to uncertainty in object properties, constraints, and coupling phenomenon with robot control systems. Humans learn motion primitives over time to effectively interact with the environment. We postulate that autonomous manipulation can be enabled by basic sets of motion primitives as well, but do not necessitate mimicking human motion primitives. Here, this work presents an approach to generalized optimal motion primitives using physics-informed neural networks. Our simulated and experimental results demonstrate that optimality is notionally maintained where the mean maximum observed final position percent error was 0.564% and the average mean error for all the trajectories was 1.53%. These results indicate that notional generalization is attained using a physics-informed neural network approach that enables near optimal real-time adaptation of primitive motion profiles.

97 MATHEMATICS AND COMPUTING

Online Bayesian State Estimation for Real-Time Monitoring of Growth Kinetics in Thin Film Synthesis

Rapid validation of newly predicted materials through autonomous synthesis requires real-time adaptive control methods that exploit physics knowledge, a capability that is lacking in most systems. Here, in this study, we demonstrate an approach to enable real-time control of thin film synthesis by combining in situ optical diagnostics with a Bayesian state estimation method. We developed a physical model for film growth and applied the direct filter (DF) method for real-time estimation of nucleation and growth rates during pulsed laser deposition (PLD). We validated the approach using simulated and experimental reflectivity data for WSe 2 growth and ultimately deployed the algorithm on an autonomous PLD system during the growth of 1T'-MoTe 2 . The DF robustly estimates growth parameters in real time at early stages of growth, down to 15% monolayer area coverage. This fusion of in situ diagnostics, data assimilation, and physical modeling opens new opportunities in adaptive control of synthesis trajectories toward desired material states.

36 MATERIALS SCIENCE

Advanced Computational Techniques for Improving Resilience of Critical Energy Infrastructure under Cyber-Physical Attacks

In this chapter, we present recent advances in improving the resilience of cyber-physical systems, especially with regards to energy systems. We provide discussions around various types of cyber-physical events that can cause disruptions and new advances in optimization, control, and reinforcement learning (RL) to deal with the challenges posed by such cyber-physical events. The presented methods range from distributed robust optimization, autonomous and coordinated control, reinforcement learning based resilient control and topology reconfiguration in Inter-System resilient control.

Nazir, Mohammad Nawaf [BATTELLE (PACIFIC NW LAB)]

Safe Physics-Informed Machine Learning for Dynamics and Control

This tutorial paper focuses on safe physics-informed machine learning in the context of dynamics and control, providing a comprehensive overview of how to integrate physical models and safety guarantees. As machine learning techniques enhance the modeling and control of complex dynamical systems, ensuring safety and stability remains a critical challenge, especially in safety-critical applications like autonomous vehicles, robotics, medical decision-making, and energy systems. We explore various approaches for embedding and ensuring safety constraints, including structural priors, Lyapunov and Control Barrier Functions, predictive control, projections, and robust optimization techniques. Additionally, we delve into methods for uncertainty quantification and safety verification, including reachability analysis and neural network verification tools, which help validate that control policies remain within safe operating bounds even in uncertain environments. The paper includes illustrative examples demonstrating the implementation aspects of safe learning frameworks that combine the strengths of data-driven approaches with the rigor of physical principles, offering a path toward the safe control of complex dynamical systems.

Drgona, Jan

Grid Resiliency with a 100% Renewable Microgrid

San Diego Gas & Electric Company (SDG&E) installed America’s first and largest utility-scale microgrid in Borrego Springs in 2013. The first generation Borrego Springs Microgrid utilized diesel generators to form and stabilize the microgrid island, with support from grid-scale batteries and local solar photovoltaic (PV) generation. In this project, SDG&E in partnership with National Renewable Energy Laboratory (NREL) demonstrated through modeling, simulation and utility field testing that blackstart and islanding of the microgrid can be led with 100% renewable, inverter based resources (IBRs), to help reduce community reliance on conventional generation resources. Through equipment upgrades, grid-forming island leader capability was transitioned to a battery IBR instead of the Borrego Springs Microgrid diesel generators. A new microgrid controller was integrated to the microgrid and programmed to control and manage multiple energy storage systems. Synchrophasor and other power quality data verified autonomous, high-speed response of the IBRs through blackstart, islanding, and load step testing. Results of project field evaluations provide distribution systems operators (DSO) with increased confidence that renewable, IBR can replace traditional generators to blackstart and island microgrids and rapidly establish stable island frequency with rapid changes in peak power demand. Importantly, the project validated the integration feasibility of a distributed energy resource management system (DERMS) controller that manages multiple grid-forming and grid-following IBRs, establishing a standard design interface to reduce the complexity of integrating new DERs in the future and supporting replication by the industry. As a result of learnings in this project, SDG&E has implemented the microgrid controller strategy at multiple other microgrid sites, thereby validating the replicability of the solution. Hardware-in-the-loop (HIL) simulations including power and controller HIL hardware — along with electromagnetic transient (EMT) simulations of Borrego Springs Microgrid —informed adjustments to inverter parameters and were important to characterize the performance of the IBRs in relevant operating conditions before deployment. The EMT and HIL simulations of islanding the entire community are important contributions in providing confidence in IBR performance prior to future islanding of the community in the field. High-fidelity EMT and/or HIL simulation of IBRs can de-risk field operations, and its relevance and importance as a tool is increasing as distribution grids and microgrids become more complex and dynamic with an increasing proportion of renewable generation, distributed energy storage, and two-way power and energy flows.

24 POWER TRANSMISSION AND DISTRIBUTION

Human Factors Considerations for Anticipated Uses of Autonomy in Advanced Reactor Technologies

The objective of this report is examine anticipated advanced reactor designs that incorporate high levels of autonomy in operation of advanced reactor systems or facilities, and to begin identifying the human factors engineering (HFE) considerations relevant to those designs. The insights gained will help identify areas where U.S. Nuclear Regulatory Commission (NRC) may need to develop or update review guidance to adequately address HFE considerations essential for safe operation. The adequacy of existing guidance, along with recommendations for updates or new guidance, may be addressed in future reports. We use the term “landscape analysis” to describe the review, analysis, and characterization of publicly available information on the anticipated future state of the industry regarding particular technologies or operational concepts. This landscape analysis considers the potential range of advanced reactor automation concepts that may be used for monitoring and controlling nuclear systems and facilities, including the potential for highly automated, near autonomous, or fully autonomous operation of advanced nuclear reactors.

99 - GENERAL AND MISCELLANEOUS

Development of Automated Atom Probe Tomography capability to study the influence of applied voltage and laser power on the final apparent composition of the analyzed specimen

This study presents the development and implementation of an autonomous Bayesian optimization (BO) framework for controlling and optimizing experimental parameters in Atom Probe Tomography (APT). Using commercial silicon needle samples as a benchmark system, we demonstrate that BO can efficiently navigate the complex parameter space of voltage and laser power to achieve target charge state ratios (specifically Si + /(Si + +Si 2+ )) with minimal experimental evaluations. Our implementation integrates Gaussian Process modeling with the CAMECA atom probe control framework, enabling autonomous adjustment of experimental conditions in real-time. Results show that the algorithm successfully converges to target ratios under different scenarios: maintaining a reference ratio, increasing the ratio (favoring Si 1+ ), and decreasing the ratio (favoring Si 2+ ). The system adapts to specimen evolution during analysis, compensating for changes in apex geometry while maintaining optimization targets. This work establishes a proof of concept for AI-driven optimization in APT, addressing the traditional challenges of manual parameter tuning and paving the way for applications to more complex materials where compositional accuracy is critical.

36 MATERIALS SCIENCE

Nuclear microreactor transient and load-following control with deep reinforcement learning

The economic feasibility of nuclear microreactors will depend on minimizing operating costs through advancements in autonomous control, especially when these microreactors are operating alongside other types of energy systems (e.g., renewable energy). This study explores the application of deep reinforcement learning (RL) for real-time drum control in microreactors, exploring performance in regard to load-following scenarios. By leveraging a point kinetics model with thermal and xenon feedback, we first establish a baseline using a single-output RL agent, then compare it against a traditional proportional–integral–derivative (PID) controller. This study demonstrates that RL controllers, including both single- and multi-agent RL (MARL) frameworks, can achieve similar or even superior load-following performance as traditional PID control across a range of load-following scenarios. In short transients, the RL agent was able to reduce the tracking error rate in comparison to PID by one half to one third. Over extended 300-minute load-following scenarios in which xenon feedback becomes a dominant factor, PID maintained better accuracy, but RL still remained within a 1% error margin despite being trained only on short-duration scenarios. This highlights RL’s strong ability to generalize and extrapolate to longer, more complex transients, affording substantial reductions in training costs and reduced overfitting. Furthermore, when control was extended to multiple drums, MARL enabled independent drum control as well as maintained reactor symmetry constraints without sacrificing performance---an objective that standard single-agent RL could not learn. We also found that, as increasing levels of Gaussian noise were added to the power measurements, the RL controllers were able to maintain lower error rates than PID, and to do so with at least 10% and upwards of 150% less control effort. These findings illustrate RL's potential for autonomous nuclear reactor control, laying the groundwork for future integration into high-fidelity simulations and experimental validation efforts.

22 - GENERAL STUDIES OF NUCLEAR REACTORS