NASA NTRS · 20020060762
Markov Tracking for Agent Coordination
Abstract
Partially observable Markov decision processes (POMDPs) axe an attractive representation for representing agent behavior, since they capture uncertainty in both the agent's state and its actions. However, finding an optimal policy for POMDPs in general is computationally difficult. In this paper we present Markov Tracking, a restricted problem of coordinating actions with an agent or process represented as a POMDP Because the actions coordinate with the agent rather than influence its behavior, the optimal solution to this problem can be computed locally and quickly. We also demonstrate the use of the technique on sequential POMDPs, which can be used to model a behavior that follows a linear, acyclic trajectory through a series of states. By imposing a "windowing" restriction that restricts the number of possible alternatives considered at any moment to a fixed size, a coordinating action can be calculated in constant time, making this amenable to coordination with complex agents.
Keep this discovery
Explore connections, maps & timelines
Washington, Richard, Lau, Sonie. 1998-01-01. Markov Tracking for Agent Coordination. https://ntrs.nasa.gov/citations/20020060762
Cite the original work for its findings. Save a collection to share your selection of sources.