RO AI LG MA SYDec 3, 2024

TAB-Fields: A Maximum Entropy Framework for Mission-Aware Adversarial Planning

Gokul Puthumanaillam, Jae Hyuk Song, Nurzhan Yesmagambet, Shinkyu Park, Melkior Ornik

arXiv:2412.02570v15.74 citationsh-index: 14Has Code

Originality Incremental advance

AI Analysis

This addresses the challenge of adversarial planning for autonomous agents in domains like robotics, though it is incremental as it builds on existing POMDP frameworks with a novel representation.

The paper tackles the problem of autonomous agents planning in adversarial scenarios where adversaries' exact policies are unknown but their mission objectives and environmental constraints are known, by developing TAB-Fields to represent adversary state distributions and integrating it with planning algorithms, resulting in superior performance in simulations and hardware experiments compared to baselines.

Autonomous agents operating in adversarial scenarios face a fundamental challenge: while they may know their adversaries' high-level objectives, such as reaching specific destinations within time constraints, the exact policies these adversaries will employ remain unknown. Traditional approaches address this challenge by treating the adversary's state as a partially observable element, leading to a formulation as a Partially Observable Markov Decision Process (POMDP). However, the induced belief-space dynamics in a POMDP require knowledge of the system's transition dynamics, which, in this case, depend on the adversary's unknown policy. Our key observation is that while an adversary's exact policy is unknown, their behavior is necessarily constrained by their mission objectives and the physical environment, allowing us to characterize the space of possible behaviors without assuming specific policies. In this paper, we develop Task-Aware Behavior Fields (TAB-Fields), a representation that captures adversary state distributions over time by computing the most unbiased probability distribution consistent with known constraints. We construct TAB-Fields by solving a constrained optimization problem that minimizes additional assumptions about adversary behavior beyond mission and environmental requirements. We integrate TAB-Fields with standard planning algorithms by introducing TAB-conditioned POMCP, an adaptation of Partially Observable Monte Carlo Planning. Through experiments in simulation with underwater robots and hardware implementations with ground robots, we demonstrate that our approach achieves superior performance compared to baselines that either assume specific adversary policies or neglect mission constraints altogether. Evaluation videos and code are available at https://tab-fields.github.io.

View on arXiv PDF Code

Similar