Search NASA⌕ Search

SEARCH · Search NASA

Results for “Maneuver design”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

A safe reinforcement learning algorithm for supervisory control of power plants

Traditional control theory-based methods require tailored engineering for each system and constant fine-tuning. In power plant control, one often needs to obtain a precise representation of the system dynamics and carefully design the control scheme accordingly. Model-free Reinforcement learning (RL) has emerged as a promising solution for control tasks due to its ability to learn from trial-and-error interactions with the environment. It eliminates the need for explicitly modeling the environment’s dynamics, which is potentially inaccurate. However, the direct imposition of state constraints in power plant control raises challenges for standard RL methods. To address this, we propose a chance-constrained RL algorithm based on Proximal Policy Optimization for supervisory control. Our method employs Lagrangian relaxation to convert the constrained optimization problem into an unconstrained objective, where trainable Lagrange multipliers enforce the state constraints. In conclusion, our approach achieves the smallest distance of violation and violation rate in a load-follow maneuver for an advanced Nuclear Power Plant design.

constrained optimization↗

Automated Construction of Artificial Lattice Structures with Designer Electronic States

Manipulating matter with a scanning tunneling microscope (STM) enables the creation of atomically defined artificial structures that host designer quantum states. However, the time-consuming nature of the manipulation process, coupled with the sensitivity of the STM tip, constrains the exploration of diverse configurations and limits the size of the designed features. In this study, we present a reinforcement learning (RL)-based framework for creating artificial structures by spatially manipulating carbon monoxide (CO) molecules on a copper substrate by using the STM tip. The automated workflow combines molecule detection and manipulation, employing deep-learning-based object detection to locate CO molecules and linear assignment algorithms to allocate these molecules to designated target sites. We initially perform molecule maneuvering based on randomized parameter sampling for sample bias, tunneling current set point, and manipulation speed. This data set is then structured into an action trajectory used to train an RL agent. The model is subsequently deployed on the STM for real-time fine-tuning of the manipulation parameters during structure construction. Our approach incorporates path-planning protocols coupled with active drift compensation to enable atomically precise fabrication of structures with significantly reduced human input while realizing larger-scale artificial lattices with the desired electronic properties. Furthermore, using our approach, we demonstrate the automated construction of an extended artificial graphene lattice and confirm the existence of a characteristic Dirac point in its electronic structure. Further challenges regarding the RL-based structural assembly scalability are discussed.

Algorithms↗

Optimizing Non-Terrestrial Hybrid RF/FSO Links With Reinforcement Learning: Navigating Through Clouds

In the pursuit of ubiquitous broadband connectivity, there has been a significant shift towards the vertical expansion of communication networks into space, particularly through the exploitation of low Earth orbit (LEO) satellite constellations, which are favored for their relatively low latency. However, this approach faces many challenges that need to be addressed, including atmospheric turbulence, high path loss, and dynamic cloud formations. High-altitude pseudo-satellites (HAPS) have emerged as promising relaying layers between LEO satellites and ground stations, enhancing coverage, latency, and direct terrestrial user connectivity. While radio frequency (RF) bands suffer from congestion and limited bandwidth, free space optical (FSO) communications offer higher data rates, but are susceptible to misalignment and weather-induced signal degradation. To address these challenges, a hybrid RF/FSO approach has been proposed to take advantage of both technologies by dynamic switching between RF and FSO based on propagation channel conditions. This paper introduces a reinforcement learning-based algorithm designed to optimize the trajectory of HAPS, maneuver around cloudy areas, and seamlessly switch between the RF and FSO communication modes to maximize the achievable capacity. The proposed approach aims to maximize system performance by intelligently adapting to environmental conditions and offering a promising solution for next-generation space communication networks.

actor-critic algorithm↗

User’s Manual for RESRAD-RDD&IND Code Version 2: Vol. 2—User’s Guide for RESRAD-RDD&IND Code

Version 2.0 of the RESRAD-RDD&IND computer code is designed to support the implementation of protective action guides (PAGs) after a nuclear emergency incident including a radiological dispersal device (RDD) and/or an improvised nuclear device (IND) incident (EPA 2017). Eight different group types, addressing various decisions, are available for selection. The RESRAD-RDD&IND code calculates radiological doses, stay times, etc., for the selected group that the user wishes to focus on. (That is, the results for all the groups are not calculated simultaneously, and the input for those other groups do not matter, although some parameter values are shared between groups.) Version 2.0 has a user-friendly interface so that the RESRAD-RDD&IND code can be used with minimal training. For example, the user can select the major characteristics of the problem-event type, source term, and decision type from the left side of the interface and then calculate the results with the default assumptions for the exposure scenarios. More in-depth analysis would include specifying site-specific exposure scenario characteristics in the right side of the interface. The procedures for data entry and results viewing are self-explanatory. This is because common window maneuvering features and text instructions were incorporated in the interface design. General and context-specific help are available to aid users entering parameter values, as well. The RESRAD-RDD&IND computer code gives the user the option to select either an RDD or IND incident for analysis. For an RDD event analysis, 11 radionuclides (Am-241, Cf-252, Cm-244, Co-60, Cs-137, Ir-192, Po-210, Pu-238, Pu-239, Ra-226, and Sr-90) are included. These 11 radionuclides are the radionuclides most likely used for an RDD. More than 90 radionuclides can be selected for an IND event analysis. Initial default concentrations are provided for 44 radionuclides for a uranium-fueled IND event. These 44 radionuclides are those that would contribute significantly to the radiation dose associated with a uranium-fueled bomb detonation. The radionuclides generated from ingrowth of these 44 initial radionuclides are also automatically included in the analysis. Pu-239, Cs-134m, Ru-105, and Rb-89 and their progeny can be selected for analysis if they are detected and their concentrations are determined. This user’s guide, which is Volume 2 of the User’s Manual for RESRAD-RDD&IND Code Version 2, provides instructions to users on how to install the RESRAD-RDD&IND code, navigate the interface, and use the various features, including those discussed above, to set up an analysis and view/print the results in text outputs. Volume 1 of the User’s Manual for RESRAD-RDD&IND Code Version 2 (Yu et al. 2026), which contains descriptions of the methodology and theoretical basis for dose modeling and the mathematical equations implemented in the code, can be accessed and viewed through the Help menu in the code or can be downloaded from the RESRAD website (https://resrad.evs.anl.gov).

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Robust cooperative control strategy for a platoon of connected and autonomous vehicles against sensor errors and control errors simultaneously in a real-world driving environment

In a real-world driving environment, a platoon of connected and autonomous vehicles (CAVs) is subject to many internal and external disturbances, resulting in uncertain vehicle dynamics. In general, the disturbances can be categorized into two types: disturbances due to vehicle sensor errors (e.g., GPS error) and disturbances due to vehicle control errors (e.g., actuator delay). In the literature, many control strategies have been proposed to improve the robustness of the CAV platoon against uncertain vehicle dynamics induced by these disturbances. However, most of these strategies only consider one type of disturbance and cannot tackle both types of disturbances simultaneously. Furthermore, they are designed to maximize the benefits of each vehicle in the platoon independently, which can deteriorate the performance of the platoon. Here, to address these problems, this study proposes a robust cooperative control (RCC) strategy to maneuver the vehicles in the platoon cooperatively to counteract the impacts of both types of disturbances. The RCC strategy is developed based on a minimax problem, where the maximization subproblem seeks to find the worst inputs for the uncertainty terms in the vehicle dynamics equation to minimize the platoon performance, while the minimization subproblem seeks to find the optimal control decisions for all subsequent vehicles to maximize the platoon performance in the worst case. To solve the minimax problem, this study proposes a globally convergent solution algorithm. It can solve the minimax problem very efficiently to enable real time deployment of the RCC strategy. Numerical application indicates that compared to the existing methods, the RCC strategy can dramatically improve the robustness of the CAV platoon against the uncertain vehicle dynamics induced by both vehicle state detection errors and vehicle control errors. Therefore, it can maneuver the CAV platoon safely and efficiently in a real-world driving environment.

33 ADVANCED PROPULSION SYSTEMS↗

Assimilating partial observation to enhance feedback control of stochastic dynamical systems

Here, in this paper, we present a novel methodology to tackle feedback optimal control problems in scenarios where the exact state of the controlled process is unknown. It integrates data assimilation techniques and optimal control solvers to manage partial observation of the state process, a common occurrence in practical scenarios. Traditional stochastic optimal control methods assume full state observation, which is often not feasible in real-world fluid dynamics control problems. Our approach underscores the significance of utilizing observational data to inform control policy design. Specifically, we introduce a kernel learning backward stochastic differential equation (SDE) filter to enhance data assimilation efficiency and propose a sample-wise stochastic optimization method within the stochastic maximum principle framework. We demonstrate the efficacy and accuracy of our method in the control of advection-diffusion-reaction flow problem and the Dubins airplane maneuvering problem with model uncertainty.

data driven↗

Integrating vehicle trajectory planning and arterial traffic management to facilitate eco-approach and departure deployment

Eco-approach and departure (EAD) enable continuous vehicle motion in urban signalized corridors. Since such a motion can extend to the EAD vehicles’ followers, it makes EAD a promising technology to benefit the traffic flow where automated vehicles and conventional vehicles coexist. Most existing EAD studies envision an ideal setting that neglects real-world operational conditions such as lane changes, multi-movement intersection configuration, partially automated fleet, and/or limited traffic state awareness. This study aims to fill the gap by designing an EAD algorithm considering real-world traffic operation constraints. The proposed algorithm uses a model predictive controller to minimize vehicle speed reduction and variation based on the real-time traffic signal control plan and measured queues at the intersection. The required inputs are readily available at many modern intersections. We observed that the proposed controller’s performance might degrade because of lane-changing maneuvers and lead-left turn traffic signals. These observations motivated our development of a lane change management strategy and a signal control implementation strategy to facilitate the EAD implementation. The lane change management strategies separate the EAD operations and lane-changing maneuvers in time and space. The signal control implementation strategy applies lag-left turn signals to enable EAD operation for both the through and left-turn vehicles. Compared to the non-EAD case, our EAD approach produces 2.5% to 7.8% energy savings while keeping similar intersection mobility. Notably, this approach brings about 2.5% to 3.6% energy savings in a 2% CAV case. This result demonstrates the feasibility of deploying EAD at low connected automated vehicle penetration rates.

Arterial corridor management↗

Coupled Aerodynamic and Hydrodynamic Hybrid Simulation of Floating Offshore Wind Turbines

The development and innovation of floating offshore wind energy in the U.S. requires detailed high-fidelity observations and measurements of turbine and platform loading due to wind, waves, and currents. However, full-scale and quasi-full-scale experiments require significant financial and temporal investments for construction, experimental testing, and long-term field campaigns. To support the commercial advancement of the offshore wind energy industry, specialized wind tunnel and wave basin experimental facilities are critical to be able to test FOWT designs at small scale under controlled conditions prior to full-scale deployment. Oregon State University (OSU) is internationally known as a leader in water and energy research, development, and testing. The O.H. Hinsdale Wave Research Laboratory (HWRL) and the Wallace Energy Systems and Renewables Facility (WESRF) at OSU have extensive experience building, modeling, monitoring, controlling, and actuating scaled systems. Experiments on wave-structure interaction have been performed at the HWRL since its establishment in 1972. Studies have included the interaction of waves with coastal structures (breakwaters, seawalls, buildings, cylinders, bridges, fixed foundations of offshore wind turbines, etc.) and with floating structures (e.g., wave energy converters, maneuvering of vessels, etc.). Hinsdale is actively used by marine energy technology developers, both for private testing and OSU-collaborative research projects. However, despite the availability of several large-scale facilities for hydrodynamic testing (at OSU and elsewhere in the U.S.), existing experimental laboratories are generally limited in their ability to accurately generate combined wind and wave conditions. The simulation of both wind and waves in experimental testing is complicated due to a number of constraints, including: [i] incompatible similitude laws governing the wind and waves for scaled experiments, [ii] producing accurate wind over a large enough control volume via fans, and [iii] generating wind that reasonably represents the atmospheric boundary layer in existing wave basins/flumes. Hence, physical test data providing insight into the simultaneous wave- and wind-structure response of floating offshore wind components can be difficult to generate. Given the aforementioned challenges in classic hydrodynamic experiments, the motivation of this project is to establish a real-time hybrid simulation (RTHS) approach that can apply aero- and hydro-dynamic loading by augmenting wave-only experimental facilities with virtual aerodynamic forces through numerical models representing the remaining dynamic forces. RTHS is a physical-numerical approach that partitions a prototype system into physical and numerical sub-assemblies that interact with each other through actuators and sensors in real time. In coupling physical and numerical models, the hybrid simulation approach applied herein is ideal for problems with: (1) structures subjected to different scaling laws, such as floating offshore wind turbines subjected to combined aero/hydro-dynamic loading, (2) structures that are too large or complex to be tested entirely in a laboratory setting, such as deep-water mooring applications, and (3) component testing, where the behavior of a portion of the assembly is uncertain but still interacts with other portions of the structure, such as testing the fatigue life of turbine blades. Few U.S. experimental facilities are able to test simultaneous aero- and hydro-dynamic loading and none can accurately produce aero/hydro-dynamic response on scaled FOWT models due to conflicting similitude laws between the wind (commonly Reynolds) and the waves (commonly Froude). To aid in accelerating the development of the U.S. floating offshore industry, there is a significant need to develop a flexible, modular framework that can expand the capacities of existing wave-only laboratories. The project goal is to demonstrate a hydrodynamic real-time hybrid simulation (hydro-RTHS) framework that couples numerical wind and physical waves acting on a FOWT, thus representing simultaneous aero/hydro-dynamic loading. The FOWT is partitioned into a full-scale numerical sub-assembly associated with the aerodynamics and a model-scale physical sub-assembly associated with the hydrodynamics. The numerical-physical partition associated with hydro-RTHS mitigates scaling constraints by supplying different scaling laws to the physical and numerical sub-assemblies. Herein, length, force, and time are scaled and exchanged between the sub-assemblies using Froude scaling to represent the open-channel flow in the physical sub-assembly. Other similitude laws could also be utilized depending on the problem definition. It is envisioned that the ability to model FOWTs under waves and wind, with mitigation of similitude distortions, would result in reduced development costs (currently, FOWT concept development is performed with full-size pro- totypes at enormous expense and risk) and increase the reliability of the FOWT industry (since extreme wave and wind conditions and contingency events can be tested safely in a controlled environment).

16 TIDAL AND WAVE POWER↗

Effect of Surface Roughness on Dynamic Stall in Pitching Motion

Dynamic stall plays a critical role in determining the performance and stability of a wide range of fluid-dynamics systems in various engineering applications. This unsteady aerodynamic phenomenon is particularly significant for maneuvering aircraft wings, jet aircraft subjected to gust encounters, helicopter rotor blades, and wind turbine blades [1–3]. The prediction of the dynamic stall vortex (DSV) is challenging due to factors such as unsteady aerodynamics, three-dimensional (3-D) effects, turbulence and flow separation, incoming gust, and surface impact effects [4–6]. Hence, advanced computational techniques and modeling approaches in computational fluid dynamics (CFD) would be required to enhance the accuracy and reliability of DSV predictions in dynamic stall scenarios. In the past, Batther and Lee [7] employed delayed detached eddy simulations (DDES) to understand the flow physics associated with the onset of dynamic stall. Their approach demonstrated that DDES achieves results comparable to those obtained from large-eddy simulations at a reduced computational cost. In another study, Khalifa et al. [8] examined the 3-D aspects of dynamic stall on a NACA 0012 airfoil using DES solvers. In conclusion, the findings underlined the superiority of 3-D simulations over two-dimensional approaches, particularly in predicting the lift coefficient values and capturing dynamic stall stages more precisely.

97 MATHEMATICS AND COMPUTING↗